dmitripikus/llm-d-router
Inference scheduler for llm-d
Inference scheduler for llm-d
llm-d setup with coordinator performance analysis
Achieve state of the art inference performance with modern accelerators on Kubernetes
llm-d benchmark scripts and tooling
Gateway API Inference Extension
GenAI inference performance benchmarking tool
Systematic and comprehensive benchmarks for LLM systems.
KubeStellar - a flexible solution for challenges associated with multi-cluster configuration management for edge, multi-cloud, and hybrid cloud