Showing 36 of 225 projects
An open-source study on neural question generation using transformers, providing simplified training and inference pipelines.
An open-source Java framework for rapid development of machine learning and statistical applications with large dataset support.
A Rust library for natural language detection using trigram models, focusing on simplicity and performance.
A curated list of awesome resources, libraries, and tools for natural language processing (NLP) in Ruby.
A curated list of awesome resources, libraries, and tools for natural language processing (NLP) in Ruby.
A GPT-2 variant that generates plausible fake words, definitions, and usage examples from scratch.
A pretrained modeling library for Keras 3 offering simple, flexible, and fast access to models for text, image, and audio tasks.
A TensorFlow implementation of QANet for machine reading comprehension on the SQuAD dataset.
Visually build full-featured chatbots for Telegram, Facebook Messenger, Viber, Twilio, and Slack using Node-RED with minimal coding.
A flexible Python framework for developing, training, and evaluating conversational AI agents in single or multi-agent environments.
A Recurrent Neural Network library for Torch7's nn, providing RNN, LSTM, GRU, and other sequence modeling modules.
A curated list of open-source and commercial tools for labeling and managing datasets across images, audio, time series, and text.
A Python natural language processing library for pre-modern languages like Latin, Ancient Greek, and Sanskrit.
Generate datasets for AI chatbots, NLP tasks, NER, and text classification using a simple domain-specific language.
An open-source suite featuring financial large language models (FinMA), instruction datasets (FIT), and evaluation benchmarks (FinBen) for financial AI.
Catalyst is a high-performance C# NLP library inspired by spaCy, offering pre-trained models, entity recognition, and embedding training.
A minimal 200-line implementation of a sequence-to-sequence chatbot using TensorLayer and TensorFlow.
A Java API for Natural Language Generation that handles morphological realization, text generation, and basic aggregation.
A no-code natural language generation platform that transforms structured data into varied textual descriptions.
TensorFlow implementation of character-aware neural language models using CNN, highway networks, and LSTM.
A curated list of resources for Question Answering (QA), covering machine learning, deep learning, datasets, and research.
A pre-trained BERT model designed for DNA sequence analysis, enabling genome understanding tasks like classification and motif discovery.
A tool for automatically annotating mentions of DBpedia resources in text, linking entities to their global identifiers.
A Python tool that uses GPT-3.5 to read, summarize, and answer questions about academic PDF papers locally.
A modern C++ toolkit for text retrieval and analysis, featuring indexing, ranking, topic modeling, classification, and language models.
A natural language detection library for Go that identifies 84 languages and scripts with no external dependencies.
A fast, robust Python library to detect offensive language in text using a machine learning model.
A cookiecutter template for deploying spaCy NLP models as FastAPI services compatible with Azure Search Custom Skills.
A comprehensive Go library for string comparison and edit distance algorithms, including Levenshtein, LCS, Hamming, Jaro-Winkler, and Cosine similarity.
Generate Word2Vec vectors for DBpedia entities from Wikipedia dumps, linking words and topics to structured knowledge.
TensorFlow implementation of R-Net for machine reading comprehension on the SQuAD dataset.
A curated list of open-access resources and tools for Natural Language Processing (NLP) focused on the German language.
A Go library implementing word embedding models (Word2Vec, GloVe, LexVec) from scratch with CLI and SDK.
A Python library for translating between 200 languages using Hugging Face transformer models like mBART-50, m2m100, and NLLB-200.
A curated list of resources dedicated to Natural Language Generation (NLG), including datasets, libraries, tools, and research.
A curated list of resources dedicated to Natural Language Generation (NLG), including datasets, libraries, tools, and research.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.