Repositories
ys2025-AI/TransferQueue
An asynchronous streaming data management module for efficient post-training.
ys2025-AI/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
ys2025-AI/speculators
A unified library for building, evaluating, and storing speculative decoding algorithms for LLM inference in vLLM
ys2025-AI/mcore-bridge
MCore-Bridge: Providing Megatron-Core model definitions for state-of-the-art large models and making Megatron training as simple as Transformers — with support for 300+ large language models (Qwen3-Next, GLM-5.1, Deepseek-V4, MiniMax-2.7, ...) and 200+ multimodal large models (Qwen3.5, Qwen3-Omni, Gemma4, ...).
ys2025-AI/flash-linear-attention
🚀 Efficient implementations for emerging model architectures
ys2025-AI/verl
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
ys2025-AI/ms-swift
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-R1, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
ys2025-AI/twinkle
Twinkle✨: Training workbench to make your model glow.
ys2025-AI/vllm-ascend
Community maintained hardware plugin for vLLM on Ascend