#tantivy #bridge #adapter #tokenizer #jieba #jieba-rs

tantivy-jieba

A library that bridges between tantivy and jieba-rs

12 releases (breaking)

0.11.0 Apr 28, 2024
0.10.0 Oct 13, 2023
0.9.0 Jul 3, 2023
0.7.0 Jan 5, 2023
0.1.1 Feb 13, 2019

#181 in Database interfaces

Download history 25/week @ 2024-03-14 1704/week @ 2024-03-21 2443/week @ 2024-03-28 1433/week @ 2024-04-04 1480/week @ 2024-04-11 1650/week @ 2024-04-18 1864/week @ 2024-04-25 1376/week @ 2024-05-02 3084/week @ 2024-05-09 2232/week @ 2024-05-16 2004/week @ 2024-05-23 2053/week @ 2024-05-30 1846/week @ 2024-06-06 1311/week @ 2024-06-13 2148/week @ 2024-06-20 1965/week @ 2024-06-27

7,730 downloads per month

MIT license

8KB
104 lines

tantivy-jieba

Crates.io version docs.rs Changelog FOSSA Status

An adapter that bridges between tantivy and jieba-rs.

Usage

Add dependency tantivy-jieba to your Cargo.toml.

Example

use tantivy::tokenizer::*;
let mut tokenizer = tantivy_jieba::JiebaTokenizer {};
let mut token_stream = tokenizer.token_stream("测试");
assert_eq!(token_stream.next().unwrap().text, "测试");
assert!(token_stream.next().is_none());

Register tantivy tokenizer

use tantivy::schema::Schema;
use tantivy::tokenizer::*;
use tantivy::Index;
let tokenizer = tantivy_jieba::JiebaTokenizer {};
let index = Index::create_in_ram(schema);
index.tokenizers()
     .register("jieba", tokenizer);

License

FOSSA Status

Dependencies

~9.5MB
~95K SLoC