jwills/mrjob
Run MapReduce jobs on Hadoop or Amazon Web Services
A prototype of Hive UDFs/UDTFs that execute nested SQL queries within rows.
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
Run MapReduce jobs on Hadoop or Amazon Web Services
Word count done with the Scrunch (Apache Crunch for Scala) library.
A Scala API for Cascading
Hive UDF's for the data warehouse