Showing 36 of 392 projects
Simple sentiment analysis for Elixir based on AFINN-165 with emoji, booster, and negator support.
A Julia package for loading pretrained word embeddings like Word2Vec, FastText, and GloVe.
A tagger, lemmatizer, morphological analyzer, and dependency parser for Dutch using memory-based NLP modules.
A collection of Node-RED nodes to integrate IBM Watson AI services like speech, language, and conversation into applications.
A biomedical text corpus with 97 full-text articles annotated for concepts, coreferences, and structural elements.
Extract dates, times, emails, phone numbers, and other common patterns from text using pre-built regular expressions.
A Ruby gem for filtering stopwords from text with built-in support for multiple languages via Snowball lists.
A Docker-based speech recognition model that converts short English WAV audio files into text using Mozilla's DeepSpeech.
A Go package for n-gram based text categorization and language detection with UTF-8 support.
A rule-based Unicode tokenizer that separates words from punctuation and splits sentences for NLP preprocessing.
A flexible and general-purpose ngrams library written in Ruby, supporting various gram types, vocabulary models, and text analysis.
A natural language date and time parser for Common Lisp, inspired by Ruby's Chronic.
A Ruby wrapper for the spaCy NLP library via PyCall, enabling tokenization, POS tagging, NER, and OpenAI integration.
Go SDK for interacting with IBM Watson AI services, providing authentication, API clients, and utilities.
An Elixir natural language processor for tokenization, counting, and string similarity analysis.
A model-driven, rule-based natural language understanding system for high-precision information extraction at scale.
A Python library using NLP and AI to help psychologists and social scientists harmonize questionnaire items across different languages and formats.
A CCG parser implementing all combinators with parsing to logical form and parameter estimation for probabilistic CCG.
A Go implementation of the MMSEG Chinese word segmentation algorithm for text processing.
An AI-powered command-line interface for database management with natural language queries, SQL optimization, and diagnostic tools.
A TensorFlow wrapper library for character-level and word-level text generation using recurrent neural networks.
A Scala library for semantic parsing using Combinatory Categorial Grammar (CCG) to translate natural language into formal representations.
A Go library for detecting the natural language of unicode text, supporting over 60 languages.
A Julia package providing language detection, script detection, and linguistic word lists for multiple human languages.
A Go package for reading word2vec vectors and performing similarity searches and analogies.
A pre-trained BERT-based model for detecting positive or negative sentiment in short text fragments.
A BERT-based model that detects six types of toxicity in text comments, deployable as a Docker container.
A Go package providing English, German, and Dutch stemmers for natural language processing.
A Go port of VADER sentiment analysis for scoring text sentiment (positive, negative, neutral, compound).
A text mining and natural language processing API for the Elixir programming language.
A library for parsing, scheduling, and formatting repeating events from natural language expressions.
A fast and accurate rule-based sentence segmentation tool for Ruby that uses placeholder replacement.
A Playwright extension that validates web page appearance using natural language prompts instead of flaky CSS selectors.
A Go library for inflecting Russian first names, middle names, and last names into grammatical cases.
A natural language shell assistant that turns plain English prompts into executable shell commands with safety and explanation.
Ruby FFI bindings for the Hunspell spell-checking library.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.