jieba-jruby
Chinese tokenizer library
Provides a Ruby port of the popular Chinese language processing library Jieba
jieba-analysis(结巴分词) for jRuby
8 stars
3 watching
1 forks
Language: Ruby
last commit: about 12 years agoRelated projects:
| Repository | Description | Stars |
|---|---|---|
| A Ruby port of a Japanese text tokenization algorithm | 21 | |
| Provides a JRuby wrapper around the Weka Java library for machine learning and data mining in Ruby. | 65 | |
| Provides a JRuby-based backend for Ruby Debugger | 15 | |
| An Android implementation of the Chinese word segmentation algorithm jieba, optimized for fast initialization and tokenization | 153 | |
| An implementation of a feedforward neural network toolkit for JRuby | 60 | |
| A Ruby-based kernel for interactive computing environments like Jupyter Notebooks | 902 | |
| A tokenizer based on dictionary and Bigram language models for text segmentation in Chinese | 21 | |
| A Ruby-based library for splitting written text into tokens for natural language processing tasks. | 46 | |
| A Typst programming language implementation of Ruby. | 20 | |
| A Ruby library providing a standardized interface to interact with the CITA network. | 3 | |
| Provides an official Ruby client for the MonkeyLearn API to build and consume machine learning models for language processing from Ruby apps. | 80 | |
| A Ruby library that tokenizes input and provides various statistical measures about the tokens | 159 | |
| A Ruby implementation of common machine learning algorithms and techniques | 13 | |
| A Ruby library to interact with the Gengo API for translation and management tasks | 21 | |
| A JRuby client for RabbitMQ messaging system | 96 |