solrwayback

Web archiver

A search interface and archival tool for browsing historical web pages

A search interface and wayback machine for the UKWA Solr based warc-indexer framework.

GitHub

102 stars
24 watching
21 forks
Language: Java
last commit: almost 2 years ago
Linked from 1 awesome list


Backlinks from these awesome lists:

Related projects:

RepositoryDescriptionStars
wabarc/waybackA tool for capturing and preserving web content and making it accessible in the future.1,839
ukwa/webarchive-discoveryTools for indexing and discovering archived web content117
akamhy/waybackpyAn API interface and command-line tool for interacting with the Wayback Machine's web archiving service489
jarofghosts/memento-clientProvides a simple JavaScript interface to access historical web pages via the Wayback Machine14
oduwsdl/ipwbA system for dispersing and replaying archived web content using peer-to-peer technology.617
wabarc/playbackReplays archived webpages from the Wayback Machine8
nla/outbackcdxA RocksDB-based server for managing and replicating capture indexes used in web archiving33
iipc/openwaybackA Java-based tool for recording and replaying web pages from archives.487
wabarc/cairnA tool for archiving web pages as single HTML files45
ukwa/shineA web archive exploration UI built on top of the Solr search engine and warc-discovery indexer.43
machawk1/wailA graphical user interface layer for preserving and replaying web pages using multiple archiving tools.353
p3gleg/pwnbackGenerates a sitemap of a website using Wayback Machine225
netarchivesuite/jwatA toolkit for analyzing and extracting data from legacy web archives in a structured format suitable for further analysis or reuse3
ikreymer/webarchive-indexingTools for bulk indexing of WARC/ARC files to create a shared url index43
richardlehane/webarchiveProvides tools for reading and parsing web archive formats used in digital preservation.20