HAOCHENYE/lmdeploy-fork
LMDeploy is a toolkit for compressing, deploying, and serving LLMs.
Meow, meow!
LMDeploy is a toolkit for compressing, deploying, and serving LLMs.
🚀 Efficient implementations for emerging model architectures
test gh stack
Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.
AdaptiveGEMM: FP8 GEMM with Adaptation to Various Lengths of Group M
Dingo: A Comprehensive Data Quality Evaluation Tool
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
An efficient, flexible and full-featured toolkit for fine-tuning LLM (InternLM2, Llama3, Phi3, Qwen, Mistral, ...)
Collection of Evaluation Metrics and Algorithms for Machine Translation
An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.
InternLM has open-sourced a 7 and 20 billion parameter base models and chat models tailored for practical scenarios and the training system.
OpenCompass is an LLM evaluation platform, supporting a wide range of models (LLaMA, LLaMa2, ChatGLM2, ChatGPT, Claude, etc) over 50+ datasets.
:ledger: The MLOps stack component for experiment tracking
MIM Installs OpenMMLab Packages
Making large AI models cheaper, faster and more accessible
yehaochen copy