Daoyuan Li

@DaoyuanLi2816 · User

GitHub profile ↗ · Compare

Greater Seattle Area272 followers29 repositories

Repositories

DaoyuanLi2816/mcp-fence

Local-first security scanner, MCP protocol inspector, dynamic fuzzer, Docker sandbox, and report generator for Model Context Protocol servers.

★ 37PythonForks 6

DaoyuanLi2816/laptop-llm-cn

中文 LLM 研究教学实验室:Sparse Attention、MLA、MoE、PPO/GRPO、RLVR、在线蒸馏与本地网页 Serving

★ 0PythonForks 0

DaoyuanLi2816/mini-verl

verl for a single consumer GPU. PPO, GRPO and on-policy distillation on NVIDIA GPUs.

★ 304PythonForks 75

DaoyuanLi2816/RepoGuardBench

Benchmarking repository-borne prompt injection attacks and lightweight defenses for local coding agents. DL4C @ ICML 2026.

★ 70PythonForks 13

DaoyuanLi2816/verl

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

★ 0Forks 0

DaoyuanLi2816/tracedistill

Distill teacher chains-of-thought into a LoRA adapter via a strict boxed-answer format contract + two-phase Train→Nudge (silver-medal NVIDIA Nemotron reasoning recipe, as a tested library).

★ 85PythonForks 17

DaoyuanLi2816/pairjudge

Pairwise LLM judges (A/B/tie): budget-aware multi-turn packing, position-bias correction, pseudo-label distillation. Generalized from the 4th-place (gold) solution to Kaggle LMSYS Chatbot Arena.

★ 170PythonForks 12

DaoyuanLi2816/molgen

Lightweight toolkit for de novo molecular generation: SMILES & SELFIES tokenizers, CharRNN / MolGPT / VAE models, training, sampling, and MOSES-style metrics.

★ 15PythonForks 4

DaoyuanLi2816/labelbank

Retrieve + rerank over a closed label bank: LLM bi-encoders with self-mined hard negatives and a generative listwise reranker. Generalized from a silver-medal solution to Kaggle Eedi — Mining Misconceptions in Mathematics.

★ 20PythonForks 2

DaoyuanLi2816/deer-flow

An open-source long-horizon SuperAgent harness that researches, codes, and creates. With the help of sandboxes, memories, tools, skill, subagents and message gateway, it handles different levels of tasks that could take minutes to hours.

★ 0Forks 0

DaoyuanLi2816/Dolphin

The official repo for “Dolphin: Document Image Parsing via Heterogeneous Anchor Prompting”, ACL, 2025.

★ 0Forks 0

DaoyuanLi2816/DeepSpec

DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms

★ 0Forks 0

DaoyuanLi2816/accelerate

🚀 A simple way to launch, train, and use PyTorch models on almost any device and distributed configuration, automatic mixed precision (including fp8), and easy-to-configure FSDP and DeepSpeed support

★ 0Forks 0

DaoyuanLi2816/datasets

🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools

★ 0PythonForks 0

DaoyuanLi2816/trl

Train transformer language models with reinforcement learning.

★ 0PythonForks 0

DaoyuanLi2816/peft

🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.

★ 0PythonForks 0

DaoyuanLi2816/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 0Forks 0

DaoyuanLi2816/litellm

Python SDK, Proxy Server (AI Gateway) to call 100+ LLM APIs in OpenAI (or native) format, with cost tracking, guardrails, loadbalancing and logging. [Bedrock, Azure, OpenAI, VertexAI, Cohere, Anthropic, Sagemaker, HuggingFace, VLLM, NVIDIA NIM]

★ 0Forks 0

DaoyuanLi2816/openclaw

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

★ 0Forks 0