JacobHelwig/flashinfer
FlashInfer: Kernel Library for LLM Serving
Agentic RL
FlashInfer: Kernel Library for LLM Serving
List of companies offering Machine learning and Data Science internships
verl: Volcano Engine Reinforcement Learning for LLMs
Build resilient agents.
Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.
SDAR (Synergy of Diffusion and AutoRegression), a large diffusion language model(1.7B, 4B, 8B, 30B)
Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"
A high-throughput and memory-efficient inference and serving engine for LLMs
Async RL Training at Scale
Source code to accompany research paper on training multi token prediction language models using self-distillation.
dLLM: Simple Diffusion Language Modeling
Ideas for projects related to Tinker
Official Implementation of LaViDa: :A Large Diffusion Language Model for Multimodal Understanding
A framework for few-shot evaluation of language models.
Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL
A covariate-dependent approach to Gaussian graphical modeling in R
Student version of Assignment 1 for Stanford CS336 - Language Modeling From Scratch
💯 Curated coding interview preparation materials for busy software engineers
Tensors and Dynamic neural networks in Python with strong GPU acceleration
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch and FLAX.
Use Fourier transform to learn operators in differential equations.
Inference code for LLaMA models