wendyran/datafu-1
Hadoop library for large-scale data processing, now an Apache Incubator project
Public repository discovered through GitHub real-time crawl.
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
Hadoop library for large-scale data processing, now an Apache Incubator project
Public repository.
GraphX development repository (which will eventually be merged into Apache Spark)
Sina Weibo Data Crawler By API