C1rN09/xtuner
An efficient, flexible and full-featured toolkit for fine-tuning LLM (InternLM2, Llama3, Phi3, Qwen, Mistral, ...)
An efficient, flexible and full-featured toolkit for fine-tuning LLM (InternLM2, Llama3, Phi3, Qwen, Mistral, ...)
LMDeploy is a toolkit for compressing, deploying, and serving LLMs.
An automatic evaluator for instruction-following language models. Human-validated, high-quality, cheap, and fast.
OpenCompass is an LLM evaluation platform, supporting a wide range of models (LLaMA, LLaMa2, ChatGLM2, ChatGPT, Claude, etc) over 50+ datasets.
A unified evaluation library for multiple machine learning libraries
OpenMMLab Computer Vision Foundation
TorchBench is a collection of open source benchmarks used to evaluate PyTorch performance.