zyzshishui/torch_memory_saver
Allow torch tensor memory to be released and resumed later
Allow torch tensor memory to be released and resumed later
SGLang is a fast serving framework for large language models and vision language models.
Fast and memory-efficient exact attention
super repo for rocm systems projects
From Automated Idea Factory to Realization
[DEPRECATED] Moved to ROCm/rocm-systems repo
The sglang router for miles only.
AI Tensor Engine for ROCm
A high-throughput and memory-efficient inference and serving engine for LLMs
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
A CPU+GPU Profiling library that provides access to timeline traces and hardware performance counters.
slime is a LLM post-training framework aiming at scaling RL.
verl: Volcano Engine Reinforcement Learning for LLMs
Tensors and Dynamic neural networks in Python with strong GPU acceleration
Ongoing research training transformer models at scale