pub enum Tokenizer {
Blank,
Camel,
Class,
Punct,
Segment(SegmentLanguage),
}Variants§
Blank
Camel
Class
Punct
Segment(SegmentLanguage)
Splits a span into words using a morphological dictionary.
Unlike the character-class tokenizers, which decide a boundary from the pair of characters around it, this one needs the whole span at once: the boundaries come from a lattice search over dictionary entries. It therefore runs after the character-class tokenizers have produced spans, segmenting each one further.
Trait Implementations§
impl Copy for Tokenizer
impl Eq for Tokenizer
impl StructuralPartialEq for Tokenizer
Auto Trait Implementations§
impl Freeze for Tokenizer
impl RefUnwindSafe for Tokenizer
impl Send for Tokenizer
impl Sync for Tokenizer
impl Unpin for Tokenizer
impl UnsafeUnpin for Tokenizer
impl UnwindSafe for Tokenizer
Blanket Implementations§
Source§impl<T> BorrowMut<T> for Twhere
T: ?Sized,
impl<T> BorrowMut<T> for Twhere
T: ?Sized,
Source§fn borrow_mut(&mut self) -> &mut T
fn borrow_mut(&mut self) -> &mut T
Mutably borrows from an owned value. Read more
Source§impl<T> CloneToUninit for Twhere
T: Clone,
impl<T> CloneToUninit for Twhere
T: Clone,
Source§impl<Q, K> Equivalent<K> for Q
impl<Q, K> Equivalent<K> for Q
Source§fn equivalent(&self, key: &K) -> bool
fn equivalent(&self, key: &K) -> bool
Compare self to
key and return true if they are equal.