Showing 36 of 116 projects
A Julia package for loading pretrained word embeddings like Word2Vec, FastText, and GloVe.
A flexible and general-purpose ngrams library written in Ruby, supporting various gram types, vocabulary models, and text analysis.
A Ruby wrapper for the spaCy NLP library via PyCall, enabling tokenization, POS tagging, NER, and OpenAI integration.
An Elixir natural language processor for tokenization, counting, and string similarity analysis.
A model-driven, rule-based natural language understanding system for high-precision information extraction at scale.
A Go implementation of the MMSEG Chinese word segmentation algorithm for text processing.
A Go library for detecting the natural language of unicode text, supporting over 60 languages.
A Julia package providing language detection, script detection, and linguistic word lists for multiple human languages.
A Go package providing English, German, and Dutch stemmers for natural language processing.
A text stream monitoring library that detects keywords, phrases, regexes, and complex Lucene queries in documents.
A Go port of VADER sentiment analysis for scoring text sentiment (positive, negative, neutral, compound).
A text mining and natural language processing API for the Elixir programming language.
A Go library implementing the Snowball stemming algorithm for natural language processing.
A social media assistant that analyzes tweets to provide the most valuable hashtags for increasing visibility and followers.
A Go library that parses human names into discrete components like first name, last name, and generation.
A Julia package implementing Latent Dirichlet Allocation (LDA) topic models with collapsed Gibbs sampling inference.
A Ruby gem for calculating text readability statistics, complexity metrics, and grade levels across 22 languages with high performance.
A legacy Ruby SDK for accessing AlchemyAPI's text analysis and image recognition AI services.
A multilingual Rust implementation of the RAKE algorithm for automatic keyword extraction from text.
Elixir library for sentiment analysis using the AFINN-111 word list.
High-quality text language identification using Wikipedia data, supporting 156 languages with ISO-639-1 codes.
A Ruby gem for customizable text tokenization, useful for web crawling and natural language processing.
A Go implementation of the Paice/Husk stemming algorithm for natural language processing.
A Go implementation of the Paice/Husk stemming algorithm for natural language processing.
An Elixir port of Nakatani Shuyo's natural language detection library, supporting 55 languages.
A Go client library for the Detect Language API, enabling language detection and account management.
A named entity recognition model that locates and tags entities like persons, locations, and organizations in text using a neural network.
An OCaml library providing efficient access to Unicode character properties from the Unicode character database.
A Ruby gem for performing sentiment analysis on German text using dictionary-based scoring.
An Elixir library for calculating tf-idf (term frequency–inverse document frequency) scores to identify important words in text.
GitHub Action that automatically checks text for insensitive, inconsiderate, or unequal phrasing using alex.
A simple and extensible Ruby gem for sentiment analysis with customizable analysis strategies.
A Swift playground for sentiment analysis using AFINN-165 wordlist and Emoji Sentiment Ranking.
A Ruby library for breaking words and phrases into n-grams with customizable parameters.
A Common Lisp library for automatic detection of text file encoding, delimiters, end-of-line characters, and column consistency.
A JRuby wrapper for Apache OpenNLP that provides natural language processing tools like tokenization, POS tagging, and entity extraction.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.