domain: localization
6 products, primary matches first, then adoption-weighted; health v2 shown.
| Product | Health v2 | Stars | Maturity |
|---|---|---|---|
| fxsjy/jieba Jieba is the most popular Python library for Chinese word segmentation, supporting precise, full, search-engine, and PaddlePaddle-based seg… | 23 | 35130 | maintenance |
| NLPchina/ansj_seg A Java library for Chinese word segmentation based on n-Gram+CRF+HMM, achieving ~2 million characters/second with 96%+ accuracy. It also pr… | 23 | 6517 | maintenance |
| rockyzhengwu/FoolNLTK FoolNLTK is a Python toolkit for Chinese natural language processing built on a BiLSTM model. It provides high-accuracy word segmentation, … | 32 | 1678 | maintenance |
| yongzhuo/nlp_xiaojiang A Chinese natural language processing toolkit covering retrieval-based chatbots, text classification, NER (BERT+BiLSTM+CRF), sentence embed… | 32 | 1534 | maintenance |
| buppt/ChineseNER A simple Chinese named entity recognition (NER) implementation using BiLSTM+CRF in both TensorFlow and PyTorch. It includes training script… | 32 | 1463 | maintenance |
| zhanlaoban/EDA_NLP_for_Chinese A Python implementation of the EDA (Easy Data Augmentation) paper adapted for Chinese text corpora. It augments labeled text classification… | 32 | 1382 | maintenance |
page 1 / 1