danirayan/cloudera-cookbook
Cloudera (Hadoop + Hive) chef cookbook
Discovered public repositories for danirayan in the GitHub catalog.
Cloudera (Hadoop + Hive) chef cookbook
Data collection framework for OpenTSDB
A scalable, distributed Time Series Database.
A fully asynchronous, non-blocking, thread-safe, high-performance HBase client.
Continuous Streaming SQL Queries for Flume
Flume is a distributed, reliable, and available service for efficiently collecting, aggregating, and moving large amounts of log data. It has a simple and flexible architecture based on streaming data flows. It is robust and fault tolerant with tunable reliability mechanisms and many failover and recovery mechanisms. The system is centrally managed and allows for intelligent dynamic management. It uses a simple extensible data model that allows for online analytic applications.
Open PubSubHubbub Hub that does polling for you, built on top of Tatsumaki and AnyEvent
Use Avro to store all your values in HBase instead of regular columns
Mirror of Apache Hadoop HBase
cloudera repo management tool
For tracking to builds on multiple machines running Hudson. Idea stems from the Panic Board ( http://www.panic.com/blog/2010/03/the-panic-status-board/ ).
zk-smoketest.py provides a simple smoketest client for a ZooKeeper ensemble