none0663

@none0663 ยท User

GitHub profile โ†— ยท Compare

4 followers8 repositories

Repositories

none0663/slime

slime is a LLM post-training framework for RL Scaling.

โ˜… 0PythonForks 0

none0663/openclaw

Your own personal AI assistant. Any OS. Any Platform. The lobster way. ๐Ÿฆž

โ˜… 0Forks 0

none0663/verl

veRL: Volcano Engine Reinforcement Learning for LLM

โ˜… 0PythonForks 0

none0663/mbridge

Bridge Megatron-Core to Hugging Face/Reinforcement Learning

โ˜… 0Forks 0

none0663/OpenRLHF

An Easy-to-use, Scalable and High-performance RLHF Framework (70B+ PPO Full Tuning & Iterative DPO & LoRA & RingAttention & RFT)

โ˜… 0Forks 0

none0663/PARL

A high-performance distributed training framework for Reinforcement Learning

โ˜… 0Forks 0