jieba-jruby

Chinese tokenizer library

Provides a Ruby port of the popular Chinese language processing library Jieba

jieba-analysis(结巴分词) for jRuby

GitHub

8 stars
3 watching
1 forks
Language: Ruby
last commit: about 12 years ago

Related projects:

RepositoryDescriptionStars
6/tiny_segmenterA Ruby port of a Japanese text tokenization algorithm21
paulgoetze/weka-jrubyProvides a JRuby wrapper around the Weka Java library for machine learning and data mining in Ruby.65
jruby/jruby-debugProvides a JRuby-based backend for Ruby Debugger15
452896915/jieba-androidAn Android implementation of the Chinese word segmentation algorithm jieba, optimized for fast initialization and tokenization153
jedld/brains-jrubyAn implementation of a feedforward neural network toolkit for JRuby60
sciruby/irubyA Ruby-based kernel for interactive computing environments like Jupyter Notebooks902
xujiajun/gotokenizerA tokenizer based on dictionary and Bigram language models for text segmentation in Chinese21
arbox/tokenizerA Ruby-based library for splitting written text into tokens for natural language processing tasks.46
rinmyo/ruby-typA Typst programming language implementation of Ruby.20
citahub/cita-sdk-rubyA Ruby library providing a standardized interface to interact with the CITA network.3
monkeylearn/monkeylearn-rubyProvides an official Ruby client for the MonkeyLearn API to build and consume machine learning models for language processing from Ruby apps.80
abitdodgy/words_countedA Ruby library that tokenizes input and provides various statistical measures about the tokens159
mizor/machine-learning-rubyA Ruby implementation of common machine learning algorithms and techniques13
gengo/gengo-rubyA Ruby library to interact with the Gengo API for translation and management tasks21
ruby-amqp/march_hareA JRuby client for RabbitMQ messaging system96