cloudera/hcatalog-examples
Sample code for reading and writing tables with hcatalog
Discovered public repositories for cloudera in the GitHub catalog.
Sample code for reading and writing tables with hcatalog
Example programs and scripts for accessing parquet files
Java readers/writers for Parquet columnar file formats to use with Map-Reduce
Madlib port for Cloudera Impala
Sample UDF and UDAs for Impala.
Llama - Low Latency Application MAster
Public repository.
Public repository.
Public repository.
Public repository.
Public repository.
Public repository.
Public repository.
Public repository.
Public repository.
Real-time Query for Hadoop; mirror of Apache Impala
A patched Jetty 6.1.26 for use in Hadoop
Example application for analyzing Twitter data using CDH - Flume, Oozie, Hive
Public repository.
The fast and fun way to write YARN applications.
Public repository.
Public repository.
Cloudera Manager API Client
System for performing seismic data processing on a Hadoop cluster.
An analysis of adverse drug event data using Hadoop, R, and Gephi
Hadoop for archiving email
Hive + Avro. Serde for working with Avro in Hive
Puppet module to help manage Apt
Bigtop is a project for the development of packaging and tests of the Apache Hadoop ecosystem. The primary goal of Bigtop is to build a community around the packaging and interoperability testing of Hadoop-related projects. This includes testing at various levels (packaging, platform, runtime, upgrade, etc...) developed by a community with a focus on the system as a whole, rather than individual projects.
ssh, scp and sftp for java
Mirror of Apache Whirr
Mirror of Apache Hadoop MapReduce
Mirror of Apache Hadoop HDFS
Mirror of Apache Hadoop common
Oozie - workflow engine for Hadoop
Truncates lines of text to fit within a container
Toolkit of simple scripts useful for managing Hadoop
Standalone CSS Selector Parser and Engine. An official MooTools project.
Behavior Filters for MooTools More
Auto-instantiates widgets/classes based on parsed, declarative HTML.
Open source SQL Query Assistant service for Databases/Warehouses
WE HAVE MOVED to Apache Incubator. https://cwiki.apache.org/FLUME/ . Flume is a distributed, reliable, and available service for efficiently collecting, aggregating, and moving large amounts of log data. It has a simple and flexible architecture based on streaming data flows. It is robust and fault tolerant with tunable reliability mechanisms and many failover and recovery mechanisms. The system is centrally managed and allows for intelligent dynamic management. It uses a simple extensible data model that allows for online analytic applications.
Public repository.
vector drawing for buttons, icons, widgets and all that stuff
The Clientcide Javascript Libraries
Metarepository for checking out all of the hadoop subprojects
PigLatin mode for Emacs.
cloudera repo management tool