jcrobak/boto
Python interface to Amazon Web Services
Discovered public repositories for jcrobak in the GitHub catalog.
Python interface to Amazon Web Services
A release plugin for sbt (>= 0.11.0)
An AWS SDK-backed FileSystem driver for Hadoop
Metrics module for Play2
examples of running Spark and Scalding jobs over Avro data.
A workflow plugin to help many devs work with cookbooks and environments at once
Chef Cookbook for Docker
HBase Stargate (REST API) client wrapper for Python.
Kafka protocol support in Python
python implementation of the parquet columnar file format.
Public repository.
The Parquet site.
Luigi is a Python module that helps you build complex pipelines of batch jobs. It handles dependency resolution, workflow management, visualization etc. It also comes with Hadoop support built in.
Twitter's collection of LZO and Protocol Buffer-related Hadoop, Pig, Hive, and HBase code.
Twitter common libraries for python and the JVM
A more pretty, more usable web dashboard for Apache Oozie, written in Scala.
Read/Write JSON SerDe for Apache Hive
MongoDB adapter for Hadoop
A Scala productivity framework for Hadoop.
Mirror of Apache Pig
Alternate formulae repos for Homebrew
John Langford's original release of Vowpal Wabbit -- a fast online learning algorithm
A fully asynchronous, non-blocking, thread-safe, high-performance HBase client.
Mirror of Apache Avro
Maven 2 Plugin for processing Apache Avro files. Avro is a subproject of Apache Hadoop.
Hue is a browser-based desktop interface for interacting with Hadoop. It supports a file browser, job tracker interface, cluster health monitor, and more.
Job scheduler