aJupyter/ThinkLLM
ThinkLLM:🚀 轻量、高效的大语言模型算法实现
Think, Plan, Do, and Do It Better. 🤗
ThinkLLM:🚀 轻量、高效的大语言模型算法实现
my page
aJupyter
📚 《从零开始构建智能体》——从零开始的智能体原理与实践教程
搜索、推荐、广告学习笔记,王树森老师
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)
Tuning LLMs with no tears💦; Sample Design Engineering (SDE) for more efficient downstream-tuning.
torch版本的wangshusen老师NLP课程部分
Agent Zero AI framework
一个简单的多模态RAG项目
(🚧 WIP) a course of LLM inference serving on Apple Silicon for systems engineers.
Achieve your exclusive DeepResearch.
力扣刷题记录
《EasyOffer》是针对LLM宝宝们量身打造的暑期实习Offer指南,主要记录暑期实习和秋招准备的一些常见的代码和手写记录,学习ing......有问题各位大佬随时指正
一个用于transformer llm模型学习的仓库,梳理llm原理,训练的的基本步骤及微调方法, 整理能快速学习的代码实战项目
一个快速学习deepseek v3模型以及r1强化学习的仓库,侧重与理解技术报告模型设计细节
Building DeepSeek R1 from Scratch
nanoGRPO is a lightweight implementation of Group Relative Policy Optimization (GRPO)
一个论文阅读笔记
Fully open reproduction of DeepSeek-R1
Must-read Papers on LLM Agents.
全网最全-2025年AI领域最值得关注的两百位博主和一手信息源盘点
A collection of recent papers on building autonomous agent. Two topics included: RL-based / LLM-based agents.
记录大模型相关的一些知识和方法
Awesome-RAG: Collect typical RAG papers and systems.