BruceLoveDecimal/cua
Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.
Prev Bytedance
Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.
OpenCLI reborn as an MCP-native browser runtime: Chrome-spawned host, object API + code mode, site capabilities, recon
A permissioned implementation of Ethereum supporting data privacy
Model Context Protocol Server for Mobile Automation and Scraping (iOS, Android, Emulators, Simulators and Real Devices)
Make humans and AI agents work as one team — open-source and self-hostable.
Playwright is a framework for Web Testing and Automation. It allows testing Chromium, Firefox and WebKit with a single API.
A framework for efficient model inference with omni-modality models
Raft source-available release mirror. One snapshot commit per release; pull requests are not accepted.
SGLang Omni: High-Performance Multi-Stage Pipeline Framework for Omni Models
AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI
RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
UniRL is a Framework for Unified Multimodal Model Reinforcement Learning
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.
SGLang is a high-performance serving framework for large language models and multimodal models.
vLLM Quantization plugin for GGUF
Unofficial source-oriented reconstruction and extension of Grok Bot 0.18.0 for macOS
Apache Maka (Incubating) is a local-first AI agent workspace. Model messages, tool calls, tool results, permission decisions, and termination events are recorded as an append-only log.
Domain-specific language designed to streamline the development of high-performance GPU/CPU/Accelerators kernels
slime is an LLM post-training framework for RL Scaling.
A high-throughput and memory-efficient inference and serving engine for LLMs
Multimodal RL training framework for diffusion & omni models
High-performance RL post-training infrastructure. Designed to achieve bitwise operator-level train-inference consistency across heterogeneous engines and extreme memory efficiency for GRPO, PPO, etc.
Agent-Oriented Github
Personal-Model First Self Evolving AI Agent 🐘
TokenSpeed is a speed-of-light LLM inference engine.
System Level Intelligent Router for Mixture-of-Models at Cloud, Data Center and Edge