Showing 36 of 392 projects
A Go library implementing the Snowball stemming algorithm for natural language processing.
A high-performance Porter2 stemmer implementation using finite state machines for suffix comparison.
A simple tokenizer in Ruby for NLP tasks.
Ruby SDK for integrating with IBM Watson AI services like speech-to-text, language understanding, and assistant APIs.
A Zsh plugin that converts natural language descriptions into executable shell commands using AI.
A framework for keeping biomedical text mining tools running on the latest publications from PubMed.
Elixir library for inflecting Russian first, last, and middle names into grammatical cases.
An MCP server that enables natural language conversation for analyzing spatial transcriptomics data through 60+ curated methods.
Zsh plugin that converts natural language prompts into terminal commands using GPT-4.
Zsh plugin that translates natural language comments into shell commands using Claude Code AI.
A Go library and CLI tool that extracts popular tags from HTML, Markdown, or plain text documents in multiple languages.
A Julia package implementing Latent Dirichlet Allocation (LDA) topic models with collapsed Gibbs sampling inference.
A text-to-SQL tool that converts plain English questions into read-only SQL queries, automatically analyzes results with Python, and generates charts.
Go binding (cgo wrapper) for the Snowball stemming library, providing word stem extraction for multiple languages.
A phonetic algorithm for indexing Chinese characters by sound, estimating distance between words and finding similar-sounding candidates.
A Redis-backed naive Bayesian classification system for Ruby applications.
A Ruby library implementing the Earley parsing algorithm for any context-free language, with support for ambiguous grammars and parse forests.
A native Objective-C framework for experimenting with neural networks and natural language processing on macOS and iOS.
A legacy Ruby SDK for accessing AlchemyAPI's text analysis and image recognition AI services.
An incomplete collection of academic and research presentation slides by Scott Wen-tau Yih.
A multilingual Rust implementation of the RAKE algorithm for automatic keyword extraction from text.
Parse natural language dates, times, and ranges in Go without predefined formats.
A Julia package for accessing and querying Princeton's WordNet lexical database.
Elixir library for sentiment analysis using the AFINN-111 word list.
A multilingual library for parsing natural language date strings into java.util.Date objects.
A BERT-based model that answers questions from a given text corpus, deployable as a Docker container.
A Julia package providing lazy-loading iterators for various NLP corpora with automatic data dependency management.
A Ruby implementation of Naive Bayes text classification designed as a strategy for the OmniCat classification framework.
High-quality text language identification using Wikipedia data, supporting 156 languages with ISO-639-1 codes.
A Ruby gem for customizable text tokenization, useful for web crawling and natural language processing.
A JavaScript library that expands custom utterance slots for Alexa Skills Kit Sample Utterances.
Generates text using Markov chains trained on 4chan board data via command-line tool.
A Go implementation of the Paice/Husk stemming algorithm for natural language processing.
A Go implementation of the Paice/Husk stemming algorithm for natural language processing.
A Ruby natural language parser that validates if dates/times fall within complex human-readable ranges like 'Mon-Fri 9:00-16:30'.
A German ELMo deep contextualized word representation model trained on a specialized German Wikipedia text corpus for NLP tasks.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.