LukeAVanDrie/llm-d-async
Asynchronous Processor for Inference Gateway. Orchestrator of queues
Asynchronous Processor for Inference Gateway. Orchestrator of queues
repo for CI and infrastructure required to maintain llm-d org member repos
Achieve state of the art inference performance with modern accelerators on Kubernetes
llm-d Router: The intelligent entry point for inference requests
Gateway API Inference Extension
GenAI inference performance benchmarking tool
Production-Grade Container Scheduling and Management
Cornhacks 2021 Project