mailmahee/adam
A genomics processing engine and specialized file format built using Apache Avro, Apache Spark and Parquet. Apache 2 licensed.
Cloudera Development Kit
This repository is cataloged as part of our automated global GitHub synchronization. Full telemetry, velocity snapshots, and code summaries are scheduled for continuous enrichment.
A genomics processing engine and specialized file format built using Apache Avro, Apache Spark and Parquet. Apache 2 licensed.
Scalable machine learning library for Hive/Hadoop
Next-generation web analytics processing with Scala, Spark, and Parquet.
Scripts to validate a cluster is ready for MapR Hadoop installation