tmalaska/SparkOnALog
Examples of Integrating Spark Streaming, Flume, and HBase to solve Streaming problems
The rules of tera sort say you can't compress the input and output. Well those rules are out of touch with how real use cases on hadoop.
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
Examples of Integrating Spark Streaming, Flume, and HBase to solve Streaming problems
This is a tool for testing and managing many repeatedly and large bulk loads on HBase
This is a layer on top of the Flume NettyAvroRpcClient that allows for multiple connects to a server.
A simple MR job where you can declare the number of mappers and reducer and a sleep time that they will sleep for.