Showing 25 of 25 projects
A comprehensive Python library for natural language processing, providing modules, datasets, and tutorials for NLP research and development.
An open-source Python library for creating and running psychology and neuroscience experiments with both GUI and code interfaces.
A Python natural language processing library for pre-modern languages like Latin, Ancient Greek, and Sanskrit.
PyNLPl, pronounced as 'pineapple', is a Python library for Natural Language Processing. It contains various modules useful for common, and less common, NLP tasks. PyNLPl can be used for basic tasks such as the extraction of n-grams and frequency lists, and to build simple language model. There are also more complex data types and algorithms. Moreover, there are parsers for file formats common in NLP (e.g. FoLiA/Giza/Moses/ARPA/Timbl/CQL). There are also clients to interface with various NLP specific servers. PyNLPl most notably features a very extensive library for working with FoLiA XML (Format for Linguistic Annotation).
A curated list of resources, tools, datasets, and communities for linguistics and natural language processing.
A comprehensive and extensible natural language processing toolkit for Common Lisp, supporting custom pipelines and experimentation.
An R package with GUI for computational stylistics and authorship attribution through statistical text analysis.
A Ruby natural language processor for tokenizing and analyzing text with flexible filtering and custom regex support.
A Ruby interface to the WordNet lexical database, enabling natural language processing and linguistic analysis.
A C++ and Python library for efficient extraction and analysis of n-grams, skipgrams, and flexgrams from large corpora.
A graphical syntax tree generator for linguistic research that creates publication-quality tree diagrams from bracket notation.
A Ruby gem for lemmatizing English text, converting inflected words to their base dictionary forms.
A Julia package providing language detection, script detection, and linguistic word lists for multiple human languages.
A Go library for inflecting Russian first names, middle names, and last names into grammatical cases.
A Go library implementing the Snowball stemming algorithm for natural language processing.
Go binding (cgo wrapper) for the Snowball stemming library, providing word stem extraction for multiple languages.
CGo bindings for Yandex.Mystem, providing Russian morphological analysis in Go applications.
A Julia package for accessing and querying Princeton's WordNet lexical database.
Resources and interactive learning games for understanding written Romance languages using the Seven Sieves method.
A Haxe library for linguistical analysis and natural language processing with tokenization, stemming, classification, and dictionary features.
A Java API for generating German natural language text, adapted from SimpleNLG 4.
Python binding for Morfologik, a Polish morphological analyzer and stemmer.
An Elixir implementation of the Porter 2 algorithm for stemming English words.
A command-line tool that finds Typo-Bridges—typos that could plausibly be two different intended words—to explore ambiguity and intentional inclusion in text.
Analyzes 250,000 Scrabble AI self-play games to track word usage patterns and power tile statistics.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.