Categories: None [Edit]

tiny_segmenter

https://rubygems.org/gems/tiny_segmenter
https://github.com/6/tiny_segmenter
Ruby port of TinySegmenter.js for tokenizing Japanese text. Uses a Naive Bayes model that has been trained using the RWCP corpus and optimized using L1-norm regularization. The resultant model is quite compact, yet has a 95% accuracy rate.

Total

Ranking: 7,858 of 192,137
Downloads: 538,409

Daily

Ranking: 10,675 of 192,103
Downloads: 72

Depended by

RankDownloadsName
40,18630,999nhkore
181,2111,543kanji-translator

Depends on

RankDownloadsName
81,292,456,495rake
29970,934,436rspec

Owners

#GravatarHandle
1iconpag