Categories: None [Edit]

maxixe

https://rubygems.org/gems/maxixe
https://github.com/rogerbraun/Maxixe
Maxixe is an implementation of the Tango algorithm describe in the paper "Mostly-unsupervised statistical segmentation of Japanese kanji sequences" by Ando and Lee. While the paper deals with Japanese characters, it should work on any unsegmented text given enough corpus data and a tuning of the algorithm parameters.

Total

Ranking: 92,127 of 183,147
Downloads: 8,221

Daily

Ranking: 51,220 of 183,139
Downloads: 0

Depended by

RankDownloadsName

Depends on

RankDownloadsName
25818,429,918rspec
78951,817,441text

Owners

#GravatarHandle
1iconrogerbraun