mravi/incubator-phoenix
Mirror of Apache Phoenix
Discovered public repositories for mravi in the GitHub catalog.
Mirror of Apache Phoenix
Mirror of Apache HCatalog
phoenix
A Twitter source for Flume Ng.
Hue is a Web application for interacting with Apache Hadoop. It supports a file browser, job tracker interface, Hive, Pig, Impala, Oozie, HBase, Solr, Sqoop2, ZooKeeper and more.
Distributed SQL query engine for big data
Hive UDF's for the data warehouse
Secondary Index for HBase
An import river similar to the elasticsearch mysql river
set of mr jobs on hbase
Starfish is a self-tuning system for big data analytics. Starfish builds on Hadoop while adapting to user needs and system workloads to provide good performance automatically, without any need for users to understand and manipulate the many tuning knobs in Hadoop.
The Salesforce Connector will allow to connect to the Salesforce application. Almost every operation that can be done via the Salesforce's API can be done thru this connector. This connector will also work if your Salesforce objects are customized with additional fields or even you are working with custom objects.
Public repository.
Battle-hardened Ironfan-ready big data chef cookbooks, laden with best practices and love from your friends at Infochimps
Data-Intensive Text Processing with MapReduce
WE HAVE MOVED to Apache Incubator. https://cwiki.apache.org/FLUME/ . Flume is a distributed, reliable, and available service for efficiently collecting, aggregating, and moving large amounts of log data. It has a simple and flexible architecture based on streaming data flows. It is robust and fault tolerant with tunable reliability mechanisms and many failover and recovery mechanisms. The system is centrally managed and allows for intelligent dynamic management. It uses a simple extensible data model that allows for online analytic applications.
Behemoth is an open source platform for large scale document analysis based on Apache Hadoop.
Secondary indexing for structured and unstructured data in Big Table style databases.
A collection of spouts, bolts, serializers, DSLs, and other goodies to use with Storm
HBase-based BI "OLAP-ish" solution
Hadoop library for large-scale data processing
Realtime Analytics
Storm - Esper integration experiment
Cloud9 is a MapReduce library for Hadoop developed at the University of Maryland
Set of Hadoop based tools for web analytic
Multidimensional data storage with rollups for numerical data
A fully asynchronous, non-blocking, thread-safe, high-performance HBase client.
A Java Graph Library based on Graph Theory
Real time web analytics using node.js and web sockets
Stream summarizer and cardinality estimator.
HBase Writes Distributor
This implementation of Bloom Filter is written in java, and uses a murmur hash. The class is a java generic class that requires a ToBytes object for converting your keys into byte arrays. However if you already have byte array keys that you wish to use them with the bloom filter. There are methods that except and test byte arrays with offsets and lengths so that you can reuse your byte array buffers.
Public repository.
A HBase schema manager using XML based table definition files.