Showing 6 of 258 projects
Implementation of the Partitioning Around Medoids (PAM) clustering algorithm within the H2O-3 machine learning platform.
A Hadoop/MapReduce tool that splits and partitions web archive records in (W)ARC files by MIME type and year.
A sample Spark job demonstrating analytics on Cassandra with SSL encryption for secure data processing.
An H2O-3 implementation of the Gap Statistic method for determining the optimal number of clusters in a dataset.
Import CSV files from AWS S3 into Cassandra using Apache Spark with a simple configuration-based approach.
A Java project that reads from and writes to TDengine using Apache Spark for data processing.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.