Docs
  • Guides
  • APIs (opens in a new tab)
  • Products
    Elasticsearch Observability Security Elastic Cloud
  • Reference
  • Troubleshoot
  • Release notes
  1. Docs /
  2. Reference /
  3. Elasticsearch /
  4. Text analysis components /
  5. Token filter reference

Normalization token filters

There are several token filters available which try to normalize special characters of a certain language.

Arabic
arabic_normalization
German
german_normalization
Hindi
hindi_normalization
Indic
indic_normalization
Kurdish (Sorani)
sorani_normalization
Persian
persian_normalization
Scandinavian
scandinavian_normalization, scandinavian_folding
Serbian

serbian_normalization

Previous
N-gram
Next
Pattern capture
  • View as Markdown
  • Report a docs issue
  • Edit this page
  • Learn how to contribute
Get started free
  • 14-day free trial
  • All features included
  • No setup required
Section
Reference
Docs Guides APIs (opens in a new tab) Reference Troubleshoot Release notes
Version
Reference Elasticsearch
Text analysis components
  • Overview
  • Analyzer reference
    • Fingerprint
    • Keyword
    • Language
    • Pattern
    • Simple
    • Standard
    • Stop
    • Whitespace
  • Tokenizer reference
    • Character group
    • Classic
    • Edge n-gram
    • Keyword
    • Letter
    • Lowercase
    • N-gram
    • Path hierarchy
    • Pattern
    • Simple pattern
    • Simple pattern split
    • Standard
    • Thai
    • UAX URL email
    • Whitespace
  • Token filter reference
    • Apostrophe
    • ASCII folding
    • CJK bigram
    • CJK width
    • Classic
    • Common grams
    • Conditional
    • Decimal digit
    • Delimited payload
    • Dictionary decompounder
    • Edge n-gram
    • Elision
    • Fingerprint
    • Flatten graph
    • Hunspell
    • Hyphenation decompounder
    • Keep types
    • Keep words
    • Keyword marker
    • Keyword repeat
    • KStem
    • Length
    • Limit token count
    • Lowercase
    • MinHash
    • Multiplexer
    • N-gram
    • Normalization
    • Pattern capture
    • Pattern replace
    • Phonetic
    • Porter stem
    • Predicate script
    • Remove duplicates
    • Reverse
    • Shingle
    • Snowball
    • Stemmer
    • Stemmer override
    • Stop
    • Synonym
    • Synonym graph
    • Trim
    • Truncate
    • Unique
    • Uppercase
    • Word delimiter
    • Word delimiter graph
  • Character filter reference
    • HTML strip
    • Mapping
    • Pattern replace
  • Normalizers
Elastic logo
  • Trademarks
  • Terms of Use
  • Privacy
  • Sitemap

© 2026 Elasticsearch B.V. All Rights Reserved.

This content is available in different formats for convenience only. All original licensing terms apply.

Elasticsearch is a trademark of Elasticsearch B.V., registered in the U.S. and in other countries. Apache, Apache Lucene, Apache Hadoop, Hadoop, HDFS and the yellow elephant logo are trademarks of the Apache Software Foundation in the United States and/or other countries.