MengsD/Megatron-Bridge-slime
Training library for Megatron-based models
Training library for Megatron-based models
slime is an LLM post-training framework for RL Scaling.
Trinity-RFT is a general-purpose, flexible and scalable framework designed for reinforcement fine-tuning (RFT) of large language models (LLM).
Start building LLM-empowered multi-agent applications in an easier way.
OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
Dependent jars for PAST