VincyZhang

@VincyZhang · User

GitHub profile ↗ · Compare

13 followers13 repositories

Repositories

VincyZhang/dynamo

A Datacenter Scale Distributed Inference Serving Framework

★ 0RustForks 0

VincyZhang/LMCache

LMCache: Supercharge Your LLM with the Fastest KV Cache Layer

★ 0PythonForks 0

VincyZhang/Mooncake

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

★ 0Forks 0

VincyZhang/llm-d

Achieve state of the art inference performance with modern accelerators on Kubernetes

★ 0ShellForks 1

VincyZhang/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 0PythonForks 0

VincyZhang/ucx

Unified Communication X (mailing list - https://elist.ornl.gov/mailman/listinfo/ucx-group)

★ 0CForks 0

VincyZhang/GenAIExamples

Generative AI Examples is a collection of GenAI examples such as ChatQnA, Copilot, which illustrate the pipeline capabilities of the Open Platform for Enterprise AI (OPEA) project.

★ 0Forks 0

VincyZhang/intel-extension-for-transformers

Extending Hugging Face transformers APIs for Transformer-based models and improve the productivity of inference deployment. With extremely compressed models, the toolkit can greatly improve the inference efficiency on Intel platforms.

★ 0PythonForks 0