vbajaria/presto
Distributed SQL query engine for big data
A library of examples showing how to use the Common Crawl corpus.
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
Distributed SQL query engine for big data
Utilities for Hadoop Map reduce
Content based and collaborative filtering based recommendation engine implementation on Hadoop and Storm
Mirror of Apache Flume