A&W

@Alex-Wengg · User

GitHub profile ↗ · Compare

AI Researcher

80 followers28 repositories

Repositories

Alex-Wengg/LMCache

LMCache: Supercharge Your LLM with the Fastest KV Cache Layer

★ 0Forks 0

Alex-Wengg/sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

★ 0Forks 0

Alex-Wengg/sglang-omni

SGLang-Omni is a high-performance serving framework for audio models (TTS, ASR) and unified multimodal models.

★ 0Forks 0

Alex-Wengg/production-stack

vLLM’s reference system for K8S-native cluster-wide deployment with community-driven performance optimization

★ 0Forks 0

Alex-Wengg/kserve

Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes

★ 0Forks 0

Alex-Wengg/diffusers

🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

★ 0Forks 0

Alex-Wengg/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 0Forks 0

Alex-Wengg/iree

A retargetable MLIR-based machine learning compiler and runtime toolkit.

★ 0Forks 0

Alex-Wengg/kokoro-coreml

Kokoro TTS on CoreML — 25× real-time on M4 Mac Mini, 17× on iPhone 16 Pro, ANE-optimized 7-stage pipeline.

★ 0Forks 0

Alex-Wengg/mlx-audio

A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.

★ 0PythonForks 0

Alex-Wengg/Starcraft2Bot

So this Terran bot can harvest minerals and vespene gas and build up a small army of marines and medivacs send to attack the enemy base.

★ 0PythonForks 0