rakeshnair/kafka-spark-consumer
Public repository.
Discovered public repositories for rakeshnair in the GitHub catalog.
Public repository.
simple scala command line options parsing
Mirror of Apache Spark
Mirror of Apache Ambari
Simple YARN application
Cloudera Manager API Client
Deploying apache-hadoop in a virtualized cluster as easy as 1-2-3.
Mirror of Apache Spark
Kafka River Plugin for ElasticSearch
Distributed and fault-tolerant realtime computation: stream processing, continuous computation, distributed RPC, and more
A collection of spouts, bolts, serializers, DSLs, and other goodies to use with Storm
Metamarkets Druid Data Store
Learn to use Storm!
Frozen version of JZMQ that is tested to work with Storm.
logstash - logs/event transport, processing, management, search.
A scalable, distributed Time Series Database.
Ruby wrapper for the ZooKeeper C client library
Zipkin is a distributed tracing system
WE HAVE MOVED to Apache Incubator. https://cwiki.apache.org/FLUME/ . Flume is a distributed, reliable, and available service for efficiently collecting, aggregating, and moving large amounts of log data. It has a simple and flexible architecture based on streaming data flows. It is robust and fault tolerant with tunable reliability mechanisms and many failover and recovery mechanisms. The system is centrally managed and allows for intelligent dynamic management. It uses a simple extensible data model that allows for online analytic applications.
Mirror of Apache Kafka
Fluentd event collector, Logs as JSON stream
Scribe is a server for aggregating log data streamed in real time from a large number of servers. It is designed to be scalable, extensible without client-side modification, and robust to failure of the network or any specific machine.
Splunk app for archive management, including HDFS support.
An open source clone of Amazon's Dynamo.