comaniac/latex-proj-tool
A toolset to traverse and manipulate a Latex project with multiple .tex files.
LLM systems
A toolset to traverse and manipulate a Latex project with multiple .tex files.
NeMo: a toolkit for conversational AI
A high-throughput and memory-efficient inference and serving engine for LLMs
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
Benchmark PyTorch Custom Operators
SGLang is a fast serving framework for large language models and vision language models.
A simple Google App Script that checks the status of USCIS processing cases
[ICML 2023] SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models
🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.
LLMPerf is a library for validating and benchmarking LLMs
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
Enable everyone to develop, optimize and deploy AI models natively on everyone's devices.
Microsoft Collective Communication Library
Personal Website
Hackable and optimized Transformers building blocks, supporting a composable construction.
🚀 A simple way to train and use PyTorch models with multi-GPU, TPU, mixed-precision
Enabling PyTorch on Google TPU
Statistics of Hugging Face hub models
Tensors and Dynamic neural networks in Python with strong GPU acceleration
Auto parallelization for large-scale neural networks
An automated Spark to FPGA accelerator framework.