Yan Bai

@ISEEKYAN · User

GitHub profile ↗ · Compare

NVIDIAShanghai137 followers19 repositories

Repositories

ISEEKYAN/mbridge

Bridge Megatron-Core to Hugging Face/Reinforcement Learning

★ 232PythonForks 78

ISEEKYAN/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 0PythonForks 0

ISEEKYAN/verl-recipe

A set of examples based on verl for end-to-end RL training recipes.

★ 0PythonForks 0

ISEEKYAN/verl

verl: Volcano Engine Reinforcement Learning for LLMs

★ 5PythonForks 0

ISEEKYAN/DeepGEMM

DeepGEMM: clean and efficient FP8 GEMM kernels with fine-grained scaling

★ 0Forks 0

ISEEKYAN/DeepEP

DeepEP: an efficient expert-parallel communication library

★ 0Forks 0

ISEEKYAN/k3

Megatron Lite support for moonshotai/Kimi-K3 (KDA + gated MLA hybrid attention, LatentMoE, MXFP4 weights) — external model integration example

★ 5PythonForks 0

ISEEKYAN/hy3

Standalone Tencent Hy3 support for Megatron-Lite — reference example of external model integration via register_model()

★ 2PythonForks 0

ISEEKYAN/miles

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

★ 0Forks 0

ISEEKYAN/vime

An LLM post-training framework with vLLM for RL Scaling

★ 0Forks 0

ISEEKYAN/slime

slime is an LLM post-training framework for RL Scaling.

★ 0Forks 0