Showing 36 of 624 projects
A Python package for automated univariate and bivariate data analysis and visualization to streamline machine learning workflows.
A Python package for exploring and analyzing data from your Home Assistant database.
A JupyterHub extension for publishing notebooks and apps as secure, interactive dashboards for non-technical audiences.
Archived R package for accessing open data from various government and scientific sources.
A Python library for logging ML metrics, parameters, and models in simple file formats, compatible with DVC and Git.
A command-line tool to view Jupyter notebooks directly in the terminal with customizable display options.
A Clojure library providing data-frames and arrays through Python's pandas and numpy.
A tool for data visualization and statistical analysis of threat intelligence indicator feeds to measure their quality and effectiveness.
A Python machine learning and informatics suite for analyzing, mining, and modeling chemical and materials data.
A collection of examples demonstrating how to use Comet.ml for machine learning experiment tracking across various Python frameworks.
Analysis of High Frequency Trading patterns and strategies on Bitcoin exchanges using Jupyter notebooks.
A Scala/Spark library for measuring fairness and mitigating bias in large-scale machine learning workflows.
A curated collection of academic papers, articles, and resources on credit scoring and credit risk modeling techniques.
A Python probabilistic programming framework for objective model selection in time-varying parameter time series models.
A practical guide to exploratory data analytics using Hadoop with Pig and Ruby for terabyte-scale data processing.
Julia package providing easy access to 700+ standard R datasets for data analysis and statistical learning.
Run Jupyter notebooks as REST API endpoints, enabling programmatic execution of notebook workflows.
An R package that installs packages from MRAN snapshots to ensure reproducible environments by locking package versions to a specific date.
A JRuby gem that provides Ruby-friendly access to Apache Mahout's scalable machine learning capabilities for recommendations.
A bridge library enabling Clojure to call R functions and use R objects for statistical computing and data science.
A resource and evaluation framework for benchmarking link prediction models on large-scale, heterogeneous biomedical knowledge graphs.
A scalable high-performance platform for R that enables large-scale machine learning, statistical analysis, and graph processing across clusters.
A Julia package for reproducible data setup, automating dataset downloads and management for scientific computing.
A curated list of colleges and universities worldwide offering data science degrees.
Open-source implementation of the winning solution for the 2018 Data Science Bowl Kaggle competition using PyTorch and U-Net.
A simple machine learning framework written in Swift, currently focusing on regression algorithms.
An open-source starter solution for the Kaggle Toxic Comment Classification Challenge, providing ready-to-use machine learning pipelines for detecting online harassment.
A GitHub Action to build and push Jupyter-enabled Docker images from data science repositories using repo2docker.
An open platform for hosting and participating in data science challenges focused on open science and open data.
A GitHub Action that automatically tests Jupyter notebooks from top to bottom using nbmake and pytest.
A desktop application for interactive computing with Jupyter notebooks, supporting multiple kernels and rich outputs.
An R package providing 2,260 network datasets in igraph format from diverse sources like social networks, animal interactions, and movie co-stars.
An R package that provides a bidirectional interface for calling Julia code from R and mapping objects between both languages.
A Rust DataFrame and data engineering library with PySpark/SQL-like syntax, built for business data pipelines with Microsoft stack integration.
A convenience meta-package that loads essential Julia packages for statistics with a single import.
An in-memory machine learning library for Scala with a scikit-learn-like API, built on Breeze for parallel and distributed systems.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.