gojieba

Chinese Word Segmentation

Provides a Go implementation of Chinese word segmentation algorithms

"结巴"中文分词的Golang版本

GitHub

2k stars
67 watching
303 forks
Language: Go
last commit: almost 2 years ago
Linked from 2 awesome lists


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
fudannlp/fnlpA toolkit for Chinese natural language processing tasks2,648
ibm/max-chinese-phonetic-similarity-estimatorEstimates phonetic similarity between Chinese words and suggests similar-sounding candidates35
cluebenchmark/cluecorpus2020A large-scale Chinese corpus for pre-training language models.927
burgyn/mmlib.tostringAutomatically generates a ToString method for classes with a custom attribute18
cluebenchmark/cluepretrainedmodelsProvides pre-trained models for Chinese language tasks with improved performance and smaller model sizes compared to existing models.806
soloice/chinese-character-recognitionThis project demonstrates how to build and train a convolutional neural network (CNN) to recognize Chinese characters.200
wenyan-lang/wenyanA programming language designed to resemble ancient Chinese grammar and syntax, compiling to JavaScript or other languages.19,790
hongshenghu/membership-inference-machine-learning-literatureA curated collection of papers on membership inference attacks and defenses in machine learning models.296
huangzworks/real-world-haskell-cnTranslation of an influential Haskell book into Chinese.1,562
fangyidong/json-simpleA simple toolkit for encoding and decoding JSON text in Java748
android-cn/android-jobsA comprehensive list of Android job openings in China2,200
hkust-knowcomp/jweThis is a software project that trains and evaluates word embeddings for Chinese words, characters, and fine-grained subcharacter components.99
clue-ai/chatyuanLarge language model for dialogue support in multiple languages1,903
embedding/chinese-word-vectorsProvides pre-trained vectors with various properties for downstream tasks in natural language processing11,874
isnowfy/snownlpA Python library for processing and analyzing Chinese text6,454