guoyunsky/reader
Free and open source feeds reader, including all major Google Reader features
Sub-Second Queries on Large Data
Public repository record indexed from GitHub. Explore verified star velocity metrics, source code repositories, and curated developer tool directories across the GitHubRepo ecosystem.
Free and open source feeds reader, including all major Google Reader features
Heritrix is the Internet Archive's open-source, extensible, web-scale, archival-quality web crawler project.
Daemon intended to monitor a queue to which Heritrix will submit URLs. On receipt, the URL is submitted to a webservice (currently via django-phantomjs) and stores the response, a modified HAR record, in a WARC file.
Python library for reading and writing warc files