Cody Yu

@comaniac · User

GitHub profile ↗ · Compare

LLM systems

@openaiSan Francisco, CA425 followers56 repositories

Repositories

comaniac/latex-proj-tool

A toolset to traverse and manipulate a Latex project with multiple .tex files.

★ 27PythonForks 5

comaniac/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 1PythonForks 0

comaniac/ray

Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.

★ 0PythonForks 1

comaniac/epoi

Benchmark PyTorch Custom Operators

★ 14Jupyter NotebookForks 3

comaniac/sglang

SGLang is a fast serving framework for large language models and vision language models.

★ 0PythonForks 0

comaniac/smoothquant

[ICML 2023] SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models

★ 0Forks 0

comaniac/transformers

🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.

★ 0PythonForks 0

comaniac/llmperf

LLMPerf is a library for validating and benchmarking LLMs

★ 1Forks 0

comaniac/FastChat

An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.

★ 0PythonForks 0

comaniac/mlc-llm

Enable everyone to develop, optimize and deploy AI models natively on everyone's devices.

★ 0Forks 0

comaniac/xformers

Hackable and optimized Transformers building blocks, supporting a composable construction.

★ 0Forks 0

comaniac/accelerate

🚀 A simple way to train and use PyTorch models with multi-GPU, TPU, mixed-precision

★ 0Forks 0

comaniac/pytorch

Tensors and Dynamic neural networks in Python with strong GPU acceleration

★ 0C++Forks 0

comaniac/alpa

Auto parallelization for large-scale neural networks

★ 0PythonForks 0

comaniac/S2FA

An automated Spark to FPGA accelerator framework.

★ 8JavaForks 3