CS PhD @ Carnegie Mellon
Repositories
YiyanZhai/modern-gpu-programming-for-mlsys
modern gpu programming
YiyanZhai/YiyanZhai.github.io
YiyanZhai/libCacheSim
a high performance library for building cache simulators
YiyanZhai/flashinfer-bench-starter-kit
FlashInfer Bench @ MLSys 2026: Building AI agents to write high performance GPU kernels
YiyanZhai/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
YiyanZhai/flashinfer-bench
Building the Virtuous Cycle for AI-driven LLM Systems
YiyanZhai/flashinfer
FlashInfer: Kernel Library for LLM Serving
YiyanZhai/fio-traces
FIO iolog traces converted from CloudPhysics traces
YiyanZhai/vidur
A large-scale simulation framework for LLM inference
YiyanZhai/mlc-llm
Enable everyone to develop, optimize and deploy AI models natively on everyone's devices.
YiyanZhai/llm-kernel-agent-results
YiyanZhai/web-llm
Bringing large-language models and chat to web browsers. Everything runs inside the browser with no server support.
YiyanZhai/cmu-catalyst.github.io
YiyanZhai/tlx
TLX - A Collection of Sophisticated C++ Data Structures, Algorithms, and Miscellaneous Helpers
YiyanZhai/mlc-assistant
Chat with your documents and improve your writing using large-language models within your browser.