Iceber/iouring-go
Provides easy-to-use async IO interface with io_uring
CNCF Ambassadorπ & Golden Kubestronaut π . Focused on Cloud Runtime, Multi-Cluster and WASM βοΈ βοΈ βοΈ @containerd reviewer & @clusterpedia-io founder
Provides easy-to-use async IO interface with io_uring
Kubernetes enhancements for Network Topology Aware Gang Scheduling & Autoscaling
Asynchronous Processor for Inference Gateway. Orchestrator of queues
Inference payload processor for llm-d
Next Generation Agentic Proxy for AI Agents and MCP servers
Native Image Generation for VS Code Agents
Python based inference-scheduler for Reinforcement Learning
Kubernetes controllers for fast model actuation using vLLM sleep/wake and launcher-based model swapping
A browser-based 3D director desk demo built with React, Vite, and Three.js.
C-ray β X-ray your containers. See beneath the runtime.
Distributed KV cache scheduling & offloading libraries
Achieve state of the art inference performance with modern accelerators on Kubernetes
repo for CI and infrastructure required to maintain llm-d org member repos
llm-d Router: The intelligent entry point for inference requests
helm charts for deploying models with llm-d
Build and run containers leveraging NVIDIA GPUs
Variant optimization autoscaler for distributed inference workloads
OME is a Kubernetes operator for enterprise-grade management and serving of Large Language Models (LLMs)
Backup and migrate Kubernetes applications and their persistent volumes
A Tieba-inspired community forum built with AI-assisted capabilities. In v1, users can create and join βBarsβ (ε§) to organize topics and interests, publish posts, and interact through threaded replies. AI is used to improve discovery, moderation, and content quality (without replacing human community culture).
Container and file artifact promotion tooling for the Kubernetes project