Showing 9 of 9 projects
A Python library for Chinese text segmentation, offering multiple modes, custom dictionaries, and keyword extraction.
A Go library for efficient multilingual text segmentation and NLP, supporting English, Chinese, Japanese, and more.
A high-performance Golang port of the Jieba Chinese text segmentation library.
A PHP Chinese text segmentation module offering precise, full, and search engine modes with support for Traditional Chinese and CJK languages.
A Ruby implementation of ICU using CLDR to format dates, numbers, currencies, plurals, and more with full Unicode support.
An open-source Chinese text segmentation library using CRF (Conditional Random Field) algorithm with support for pinyin segmentation and part-of-speech tagging.
A Go library for Unicode text segmentation at word boundaries as defined by Unicode Standard Annex #29.
An OCaml library implementing Unicode text segmentation algorithms for grapheme cluster, word, sentence, and line break detection.
Natural language processing algorithms implemented in pure Ruby with minimal dependencies.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.