posix4e/YCSB
Yahoo! Cloud Serving Benchmark
Discovered public repositories for posix4e in the GitHub catalog.
Yahoo! Cloud Serving Benchmark
Lightweight threads and actors for the JVM
Cascading is a feature rich API for defining and executing complex and fault tolerant data processing workflows on a Hadoop cluster.
HBase adapters for Cascading
ANSI SQL for Cascading on Apache Hadoop
Cap'n Proto in pure Java
Mirror of Apache Bigtop
DEPRECATED, please use the "freebsd" repo. It also has the SVN trunk branch of the FreeBSD src tree, for git-svn
CVE-2014-0160 mass scan on the Alexa top sites.
Capstan, a tool for building and running your application on OSv.
Jetlang provides a high performance java threading library. The library is based upon Retlang.
Distributed SQL query engine for big data
an open source whistleblower platform
Twitter's collection of LZO and Protocol Buffer-related Hadoop, Pig, Hive, and HBase code.
Java readers/writers for Parquet columnar file formats to use with Map-Reduce
protobuffer support for Parquet columnar format (merged, abandoned)
Why 0 copy protobufs is important
java serialization library, proto compiler, code generator
pluggable data layer for graphite to support hbase and other stores
Carbon is one of the components of Graphite, and is responsible for receiving metrics over the network and writing them down to disk using a storage backend.
Public repository.
very experimental: install Graphite via Ansible
Starter examples of ansible playbooks & environments
A hadoop context package for graphite that works with metrics2
Like the GangliaContext for Hadoop, sends metrics to Graphite
An import river similar to the elasticsearch mysql river
Java library and runtime system for efficient state machine replication
Elegant parsing in Java and Scala - lightweight, easy-to-use, powerful.
JSqlParser parses an SQL statement and translate it into a hierarchy of Java classes. The generated hierarchy can be navigated using the Visitor Pattern
Read - Write JSON SerDe for Apache Hive.
Stomping out capitalism, one line of code at a time
A scalable, distributed Time Series Database.
Mirror of Apache Sqoop
Public repository.