ThaiTokenizer
instead.@Deprecated public final class ThaiWordFilter extends TokenFilter
TokenFilter
that use BreakIterator
to break each
Token that is Thai into separate Token(s) for each Thai word.
Please note: Since matchVersion 3.1 on, this filter no longer lowercases non-thai text.
ThaiAnalyzer
will insert a LowerCaseFilter
before this filter
so the behaviour of the Analyzer does not change. With version 3.1, the filter handles
position increments correctly.
WARNING: this filter may not be supported by all JREs. It is known to work with Sun/Oracle and Harmony JREs. If your application needs to be fully portable, consider using ICUTokenizer instead, which uses an ICU Thai BreakIterator that will always be available.
AttributeSource.State
Modifier and Type | Field and Description |
---|---|
static boolean |
DBBI_AVAILABLE
Deprecated.
True if the JRE supports a working dictionary-based breakiterator for Thai.
|
input
DEFAULT_TOKEN_ATTRIBUTE_FACTORY
DEFAULT_ATTRIBUTE_FACTORY
Constructor and Description |
---|
ThaiWordFilter(TokenStream input)
Deprecated.
Creates a new ThaiWordFilter with the specified match version.
|
ThaiWordFilter(Version matchVersion,
TokenStream input)
Deprecated.
|
Modifier and Type | Method and Description |
---|---|
boolean |
incrementToken()
Deprecated.
|
void |
reset()
Deprecated.
|
close, end
addAttribute, addAttributeImpl, captureState, clearAttributes, cloneAttributes, copyTo, equals, getAttribute, getAttributeClassesIterator, getAttributeFactory, getAttributeImplsIterator, hasAttribute, hasAttributes, hashCode, reflectAsString, reflectWith, restoreState, toString
public static final boolean DBBI_AVAILABLE
public ThaiWordFilter(TokenStream input)
@Deprecated public ThaiWordFilter(Version matchVersion, TokenStream input)
ThaiWordFilter(TokenStream)
public boolean incrementToken() throws IOException
incrementToken
in class TokenStream
IOException
public void reset() throws IOException
reset
in class TokenFilter
IOException
Copyright © 2000-2014 Apache Software Foundation. All Rights Reserved.