RandNMR73/flash-attention
Fast and memory-efficient exact attention
Fast and memory-efficient exact attention
Simple and efficient DeepSeek V3 SFT using pipeline parallel and expert parallel, with both FP8 and BF16 trainings
A Quirky Assortment of CuTe Kernels
Tile primitives for speedy kernels
FlashInfer: Kernel Library for LLM Serving
A high-performance kernel library for LLM training
[CVPR2024 Highlight] VBench - We Evaluate Video Generation
A framework for few-shot evaluation of language models.
FastVideo is a unified framework for accelerated video generation.
gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
FlashAttention (Metal Port)
Leaderboard for CS 336 Assignment 2
Some experimental python scripts to send/stream videos through TCP and UDP sockets
Stanford CS149 -- Assignment 3
This repo is for demonstration purposes only.
This is the cheat sheet Jupyter Notebook I made for my Pandas Learn in One Video Tutorial. I basically condensed the Pandas API down into this one cheat sheet with hundreds of examples. I hope you find it useful.