Ernest Wong

@chewong · User

GitHub profile ↗ · Compare

68 followers22 repositories

Repositories

chewong/llm-d

llm-d enables high-performance distributed LLM inference on Kubernetes

★ 0MakefileForks 0

chewong/container-upstream

This project captures work in progress, and completed work for the Azure Core Container Upstream team

★ 0Forks 0

chewong/foundation

☁️♮🏛 This repo contains several documents related to the operation of the CNCF. File non-technical issues related to CNCF here.

★ 0Rich Text FormatForks 1

chewong/guidellm

Evaluate and Enhance Your LLM Deployments for Real-World Inference Needs

★ 0Forks 0

chewong/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 0PythonForks 0

chewong/AgentBaker

Agent Baker is aiming to provide a centralized, portable k8s agent node provisioning lib as well as rich support on different OS image with optimized k8s binaries.

★ 0GoForks 0

chewong/aibrix

Cost-efficient and pluggable Infrastructure components for GenAI inference

★ 0Forks 0

chewong/aks-gpu

Setup and configure nodes to support GPUs on k8s, with a focus on AKS nodes. This repo contains steps to build a container image with the Nvidia driver, and dependencies for integration.

★ 0Forks 0

chewong/lws

LeaderWorkerSet: An API for deploying a group of pods as a unit of replication

★ 0GoForks 0

chewong/jobset

JobSet: a k8s native API for distributed ML training and HPC workloads

★ 0Forks 0