Showing 36 of 225 projects
A pure Go library for fast, offline natural language detection supporting 29 languages.
A Go client library for interacting with the Wit.ai natural language processing HTTP API.
A modular NLP framework for extracting information from French clinical notes, compatible with spaCy and PyTorch.
A Ruby natural language processor for tokenizing and analyzing text with flexible filtering and custom regex support.
An open-source starter solution for the Kaggle Toxic Comment Classification Challenge, providing ready-to-use machine learning pipelines for detecting online harassment.
Interactive topic model visualization and interpretation library for Python, compatible with sklearn, Gensim, BERTopic, and Turftopic.
Japanese Natural Langauge Processing Libraries
An Elixir library for natural language and script detection using statistical analysis without AI.
A curated collection of Jupyter notebooks for digital humanities research and teaching, covering text analysis, data visualization, and more.
A Ruby interface to the WordNet lexical database, enabling natural language processing and linguistic analysis.
A lightweight Python library for building reproducible machine learning pipelines with minimal interface constraints.
A C++ and Python library for efficient extraction and analysis of n-grams, skipgrams, and flexgrams from large corpora.
A hands-on workshop introducing deep learning concepts with practical examples using neural networks, CNNs, RNNs, and autoencoders.
A Go implementation of the Rapid Automatic Keyword Extraction (RAKE) algorithm for extracting keywords from text.
A Ruby gem for lemmatizing English text, converting inflected words to their base dictionary forms.
Rust edit distance library accelerated with SIMD for fast Hamming, Levenshtein, and Damerau-Levenshtein calculations.
An application that uses IBM Watson AI services and Cloud Functions to analyze videos, extracting visual and audio insights for search and categorization.
A Python library providing German language support for TextBlob, enabling NLP tasks like tokenization, POS tagging, and sentiment analysis.
A Julia package providing high-performance, configurable tokenizers and sentence splitters for natural language processing.
A collection of tools, datasets, and approaches for building natural language interfaces to query the Web of Data.
Archived R package for accessing the Monkeylearn API for text classification and extraction.
Ruby bindings to the OpenNLP Java toolkit for natural language processing tasks like tokenization, POS tagging, and named entity recognition.
A collection of code samples demonstrating how to use Azure's Language Understanding (LUIS) service for natural language processing.
A curated collection of books covering Artificial Intelligence, Machine Learning, Deep Learning, and Transformers for students and professionals.
A tagger, lemmatizer, morphological analyzer, and dependency parser for Dutch using memory-based NLP modules.
Accurate Bayesian sentence tokenizer in Ruby.
A rule-based Unicode tokenizer that separates words from punctuation and splits sentences for NLP preprocessing.
A natural language date and time parser for Common Lisp, inspired by Ruby's Chronic.
InsNet Runs Instance-dependent Neural Networks with Padding-free Dynamic Batching.
A Ruby wrapper for the spaCy NLP library via PyCall, enabling tokenization, POS tagging, NER, and OpenAI integration.
An Elixir natural language processor for tokenization, counting, and string similarity analysis.
Saul is a declarative domain-specific language in Scala for designing flexible machine learning models with relational feature extraction.
A Python library using NLP and AI to help psychologists and social scientists harmonize questionnaire items across different languages and formats.
A TensorFlow wrapper library for character-level and word-level text generation using recurrent neural networks.
Kotlin implementations of TensorFlow Lite example Android apps for on-device machine learning.
A pre-trained BERT-based model for detecting positive or negative sentiment in short text fragments.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.