Repositories
MeganEFlynn/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
MeganEFlynn/speculators-research
MeganEFlynn/sglang
SGLang is a fast serving framework for large language models and vision language models.
MeganEFlynn/speculators
A unified library for building, evaluating, and storing speculative decoding algorithms for LLM inference in vLLM