Repositories
WWWjiahui/astraflow
Dataflow-Oriented Reinforcement Learning for (Multi-)Agentic LLMs
WWWjiahui/AReaL
Lightning-Fast RL for LLM Reasoning and Agents. Made Simple & Flexible.
WWWjiahui/WWWjiahui.github.io
WWWjiahui/mayavi
3D visualization of scientific data in Python
WWWjiahui/My-dashboard
WWWjiahui/llm-security
New ways of breaking app-integrated LLMs
WWWjiahui/Test-Toolchains
WWWjiahui/skills-communicate-using-markdown
My clone repository
WWWjiahui/skills-introduction-to-github
My clone repository
WWWjiahui/slime
slime is an LLM post-training framework for RL Scaling.
WWWjiahui/f24-lab10
F24 - Lab10 ReactJS/Intro to GUIS
WWWjiahui/f24-lab06
WWWjiahui/f24-lab05
WWWjiahui/f24-lab03
WWWjiahui/f24-lab02
17-214 Lab 2 (F24)