Showing 14 of 14 projects
A curated list of resources, tools, and services for web archiving, from acquisition and replay to analysis and community.
A curated list of resources, tools, and services for web archiving, from acquisition and replay to analysis and community.
A privacy-focused web archiving tool with an IM-style interface that captures pages to multiple archival services.
A Python package and CLI tool for interacting with the Wayback Machine's Save, CDX, and Availability APIs.
A Go tool and library for downloading URLs and files from Common Crawl and Wayback Machine web archives.
A web application for searching, browsing, and analyzing archived web content (ARC/WARC files) with a Solr backend.
A Chrome extension that integrates live web browsing with archived copies using the Memento protocol.
A collection of salvaged websites, articles, contributions, text, and documentation preserved from the web.
A collection of salvaged websites, articles, contributions, text, and documentation preserved from various sources.
A DuckDB extension to query web archive CDX APIs (Wayback Machine & Common Crawl) directly from SQL with smart query pushdown.
A command-line tool to playback archived webpages from the Wayback Machine using GitHub as a source.
A CLI tool to test URL availability and retrieve Internet Archive snapshots, with output in JSON, CSV, or BoltDB.
A command-line tool and Go package for archiving webpages to IPFS with Wayback Machine integration.
Intelligent web page comparison tool with Wayback Machine artifact cleaning, visual regression testing, and significance scoring.
Open-Awesome is built by the community, for the community. Submit a project, suggest an awesome list, or help improve the catalog on GitHub.