wucong25/verl-recipe
A set of examples based on verl for end-to-end RL training recipes.
A set of examples based on verl for end-to-end RL training recipes.
verl: Volcano Engine Reinforcement Learning for LLMs
Memory Estimator in RL
Multimodal RL training framework for diffusion & omni models
Community maintained hardware plugin for vLLM on Ascend
Ongoing research training transformer models at scale
verl Ascend specific recipe
Training library for Megatron-based models with bi-directional Hugging Face conversion capability
Bridge Megatron-Core to Hugging Face/Reinforcement Learning
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
爬取热门表情包并通过微信展示
黑马程序员 120天全栈区块链开发 开源教程
Spring Boot基础教程,Spring Boot 2.x版本连载中!!!