genewitch/ephasia
web-scraping data collection project
amazon aws/ec2, eucalyptus, and other scripts.
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
web-scraping data collection project
To build cached word lists for parsing web pages by statistical liklihood of usefulness based on grammar, spelling, word freq, and line length (number of non whitespace chars between ^. and $)
a C project to show why i dislike not having a primative boolean type.
Public repository.