dralves/calcite
Mirror of Apache Calcite
Distributed Systems, Databases & Blockchain
Mirror of Apache Calcite
Python implementation of TPC-C
Mirror of Apache Singa (Incubating)
Automatically exported from code.google.com/p/epilepsy-prediction
Mirror of Apache Drill
Automatically exported from code.google.com/p/ppapi
Automatically exported from code.google.com/p/nixysa
Automatically exported from code.google.com/p/supersonic
Peer Assessment 1 for Reproducible Research
Plotting Assignment 1 for Exploratory Data Analysis
Mirror of Apache Spark
Dataflow provides a simple, powerful model for building both batch and streaming parallel data processing pipelines.
Repository for Programming Assignment 2 for R Programming on Coursera
The Leek group guide to data sharing
Capturing JVM- and application-level metrics. So you know what's going on.
Mirror of Apache Whirr
DjangoCon 2011 Case Study
LogCabin is a distributed system that provides a small amount of highly replicated, consistent storage. It is a reliable place for other distributed systems to store their core metadata and is helpful in solving cluster management issues. LogCabin is still in early stages of development and is not yet recommended for actual use.
Brooklyn scripts for deploying and managing MapR
Appliance to generate a cloud-based cluster for Cloudera Certified Technology testing purposes.
incubating apis and providers for jclouds
Mirror of Apache Hadoop HBase
jclouds is an open source library that helps you get started in the cloud and reuse your java development skills. Our api allows you to freedom to use portable abstractions or cloud-specific features. We support many clouds including Amazon, VMWare, Azure, and Rackspace.
Brooklyn deployment and management of Cloudera Hadoop and Manager clusters