dongfeiwww/working-with-big-data
Slides, code, and supplemental materials for the LiveLesson: Working with Big Data: Infrastructure, Algorithms, and Visualizations
personal article repository.
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
Slides, code, and supplemental materials for the LiveLesson: Working with Big Data: Infrastructure, Algorithms, and Visualizations
Python clone of Spark, a MapReduce alike framework in Python
Mirror of Apache Spark
Duke is a fast and flexible deduplication engine written in Java