AI Infra & LLM Serving | Kubernetes, Rust, SGLang & vLLM | Building fast, scalable inference systems
15 followers21 repositories
Repositories
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.
★ 0RustForks 0
★ 0HTMLForks 0
SGLang is a high-performance serving framework for large language models and multimodal models.
★ 0PythonForks 0
An OpenList Strm tool
★ 0GoForks 0
Go rules for Bazel
★ 0GoForks 0
A high-throughput and memory-efficient inference and serving engine for LLMs
★ 0Forks 0
Custom Maccy fork with a redesigned clipboard popup and downloadable macOS builds
★ 0SwiftForks 0
Lock/unlock your Mac with your iPhone, Apple Watch, or any other Bluetooth LE devices
★ 0Forks 0
★ 2Forks 0
Demonstration of MobileSAM in the browser enabled through ONNX runtime web
★ 0Forks 0
Updated list of public BitTorrent trackers
★ 0Forks 0
Use Prometheus to monitor Kubernetes and applications running on Kubernetes
★ 0Forks 0
Go Media Framework
★ 0Forks 0
《设计数据密集型应用》中文翻译 《Designing Data-Intensive Application》
★ 0Forks 0
Spring Data InfluxDB
★ 0Forks 0
A command-line tool and Python library and Pytest plugin for automated testing of RESTful APIs, with a simple, concise and flexible YAML-based syntax
★ 0Forks 0
Avro support for Spark, SQL, and DataFrames
★ 0ScalaForks 0
The Jinja2 template engine
★ 0PythonForks 0
Prometheus for rails and sidekiq
★ 0RubyForks 1
★ 0HTMLForks 0
The ultimate Vim configuration: vimrc
★ 0VimLForks 0