# Guidance on language analyzers for non-native languages

**URL:** <https://discuss.elastic.co/t/guidance-on-language-analyzers-for-non-native-languages/384168>\
**Category:** Elasticsearch\
**Created:** [December 18, 2025, 2:18pm UTC](https://discuss.elastic.co/t/guidance-on-language-analyzers-for-non-native-languages/384168 "2025-12-18T14:18:38Z")\
**Posts on this page:** 2\
**Page:** 1

<div class="post-metadata">

**Author:** ![sibasish.palo](https://avatars.discourse-cdn.com/v4/letter/s/ea666f/32.png) [@sibasish.palo](https://discuss.elastic.co/u/sibasish.palo)\
**Post date:** [December 18, 2025, 2:18pm UTC](https://discuss.elastic.co/t/guidance-on-language-analyzers-for-non-native-languages/384168/1 "2025-12-18T14:18:38Z")

</div>

i am trying to add some new locale support and looking for guidance on language analyzers to be added for the locales. i have gone through [lang-analyzer](https://www.elastic.co/docs/reference/text-analysis/analysis-lang-analyzer) page to find the appropriate analyzers for the locales which i am trying to add.

Found the applicable analyzers for most of them but some of them doesn’t have native support in Elasticsearch, posting those here to get some guidance on the language analyzers.

| locale | Language | Applicable analyzers |
| --- | --- | --- |
| az\_AZ | Azerbaijani (Azerbaijan) | No native analyzers |
| sw\_KE | Swahili (Kenya) | No native analyzers |
| tl\_PH | Tagalog (Philippines) | No native analyzers |
| ne\_NP | Nepali (Nepal) | No native analyzers |
| si\_LK | Sinhala (Sri Lanka) | No native analyzers |

can you please suggest whether any of the existing language analyzers would work or i need to use standard analyzer for these or any other way to add the support.

Any help would be very helpful.  
Thank you

---

<div class="post-metadata">

**Author:** ![Itamar\_Syn-Hershko](https://sea2.discourse-cdn.com/elastic/user_avatar/discuss.elastic.co/itamar_syn-hershko/32/118620_2.png) [@Itamar\_Syn-Hershko](https://discuss.elastic.co/u/Itamar_Syn-Hershko)\
**Post date:** [December 31, 2025, 6:53am UTC](https://discuss.elastic.co/t/guidance-on-language-analyzers-for-non-native-languages/384168/2 "2025-12-31T06:53:02Z")

</div>

If no native analyzer exist, you could look for some online - often commercial companies offer custom language analyzers as open-source or a commercial product - eg [https://ey.fo/](https://ey.fo/) offer a search analyzer for Hebrew and so on.

If no language analyzer exists, neither for Lucene or Elasticsearch, for most languages the standard analyzer would do an ok job (although no morphological / stemming capabilities) and n-grams (4/5 grams) are often showing good performance in many languages , so you might want to (carefully) try that as well (but beware of the repercussions - [Don't use n-gram in Elasticsearch and OpenSearch - BigData Boutique Blog](https://bigdataboutique.com/blog/dont-use-n-gram-in-elasticsearch-and-opensearch-6f0b48) )

HTH
