mateiz/dsp
Demonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP
Demonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP
Deep Learning Pipelines for Apache Spark
Omnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents — swap harnesses without rewriting, enforce policies and sandboxing, and collaborate in real time from any device.
Mirror of Apache Spark
Open source platform for the complete machine learning lifecycle
Mirror of Apache Spark Website
Weld is a runtime and language for accelerating data analytics frameworks
starting html/css template. so much goodness baked in by default (previously named frontend pro template)
Java integration for Weld
Mirror of Apache Spark
Scalable Nucleotide Alignment Program -- a fast and accurate read aligner for high-throughput sequencing data
Performance tests for Spark, Shark, etc.
Scripts used to setup a Spark cluster on EC2
Hive on Spark
Mirror of Apache Pig
An open-source storage layer that brings scalable, ACID transactions to Apache Spark™ and big data workloads.
Koalas: Pandas API on Apache Spark