rbodkin/storm-trident-tutorial
Storm Trident examples
Discovered public repositories for rbodkin in the GitHub catalog.
Storm Trident examples
Stream summarizer and cardinality estimator.
Public repository.
Mirror of Apache Hadoop MapReduce
Mirror of Apache Hadoop HDFS
Mirror of Apache Hadoop common
Mirror of Apache Hive
Flume is a distributed, reliable, and available service for efficiently collecting, aggregating, and moving large amounts of log data. It has a simple and flexible architecture based on streaming data flows. It is robust and fault tolerant with tunable reliability mechanisms and many failover and recovery mechanisms. The system is centrally managed and allows for intelligent dynamic management. It uses a simple extensible data model that allows for online analytic applications.
Explorations relative to cloning FlumeJava
Mirror of Apache Avro
Avro Mapreduce Sample
ThinkBigAnalytics MapReduce Utilities