public final class ThaiWordFilter
extends TokenFilter

TokenFilter that use BreakIterator to break each Token that is Thai into separate Token(s) for each Thai word.

Please note: Since matchVersion 3.1 on, this filter no longer lowercases non-thai text. ThaiAnalyzer will insert a LowerCaseFilter before this filter so the behaviour of the Analyzer does not change. With version 3.1, the filter handles position increments correctly.

WARNING: this filter may not be supported by all JREs. It is known to work with Sun/Oracle and Harmony JREs. If your application needs to be fully portable, consider using ICUTokenizer instead, which uses an ICU Thai BreakIterator that will always be available.

Field Summary
static boolean DBBI_AVAILABLE
          True if the JRE supports a working dictionary-based breakiterator for Thai.
Constructor Summary
ThaiWordFilter(Version matchVersion, TokenStream input)
          Creates a new ThaiWordFilter with the specified match version.
Method Summary
 boolean incrementToken()
 void reset()
public static final boolean DBBI_AVAILABLE
True if the JRE supports a working dictionary-based breakiterator for Thai. If this is false, this filter will not work at all!

public ThaiWordFilter(Version matchVersion,
                      TokenStream input)
Creates a new ThaiWordFilter with the specified match version.

public boolean incrementToken()
                       throws IOException
public void reset()
           throws IOException
reset in class TokenFilter

