vMaroon/llm-d-router
Inference scheduler for llm-d
@llm-d core maintainer
Inference scheduler for llm-d
AiTer Optimized Model
Website for llm-d: This repository builds the website seen at llm-d.ai
Stage your PR review before it goes public.
AI and cloud-native proxy server and framework
Inference payload processor for llm-d
Lightweight coding agent that runs in your terminal
Achieve state of the art inference performance with modern accelerators on Kubernetes
A high-throughput and memory-efficient inference and serving engine for LLMs
Let my Claude talk to yours.
Gateway API Inference Extension
Distributed KV cache scheduling & offloading libraries
A personal PR-review extension.
GenAI inference performance benchmarking tool
llm-d benchmark scripts and tooling
A variety of instructions and demos for setting up
llm-d helm charts and deployment examples
Redis for LLMs
Scale from single vLLM instance to distributed vLLM deployment without changing any application code.
Cost-efficient and pluggable Infrastructure components for GenAI inference
LangChain for Go, the easiest way to write LLM-based programs in Go
KubeStellar - a flexible solution for challenges associated with multi-cluster configuration management for edge, multi-cloud, and hybrid cloud
Agent-Agnostic K8s based infrastructure
A Kubernetes application that utilizes Kubestellar mechanisms for multi-cluster workload staged placement.
Status add-on for Open Cluster Management transport