Chao Yang

@upfixer · User

GitHub profile ↗ · Compare

1 followers10 repositories

Repositories

upfixer/ome

Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton

★ 0Forks 0

upfixer/xai-cookbook

A collection of pragmatic, real-world examples guiding you from basic to advanced use of xAI's Grok APIs.

★ 0Forks 0

upfixer/genai-bench

Genai-bench is a powerful benchmark tool designed for comprehensive token-level performance evaluation of large language model (LLM) serving systems.

★ 0Forks 0

upfixer/sglang

SGLang is a fast serving framework for large language models and vision language models.

★ 0PythonForks 0