ByronHsu/Never-Blink
πBlink and lose.
RL System at Periodic Labs | SGLang, ex-xAI
πBlink and lose.
ππ Life as a git. Commit on your life.
Notes, demos, benchmarks, and other things I publish.
The sglang router for miles only.
108-1 NTUEE CA program assignment
Interactive theoretical Kimi-K3 inference roofline calculator for H200, B300, and GB300
Enhancing ScatterMoE
Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.
Self-Supervised Deep Learning for Fisheye Image Rectification (pytorch implementation).
Bindings for RDMA ibverbs through rdma-core
π¦ π Building FlyteGPT on Flyte with LangChain
C++ implementation of FRAIGs. Won the 1st place in 2018 Cadence-sponsored contest in NTU DSnP.
π π π Viz.js Graphviz - An Elegant Visualizer for And-Inverter Graph
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
PyTorch native quantization and sparsity for training and inference
Fast CUDA matrix multiplication from scratch
Follow the steps-by-steps tutorial on Udemy course - "Java - ambitious start. Create a real web app!" to build a mini project
Submarine is Cloud Native Machine Learning Platform.
FlashInfer: Kernel Library for LLM Serving
Learning python
Train transformer language models with reinforcement learning.
Materials for learning SGLang
Make PyTorch models up to 40% faster! Thunder is a source to source compiler for PyTorch. It enables using different hardware executors at once; across one or thousands of GPUs.
SGLang is a fast serving framework for large language models and vision language models.
Official inference repo for FLUX.1 models
Development repository for the Triton language and compiler
Material for cuda-mode lectures
Efficiently Fine-Tune 100+ LLMs in WebUI (ACL 2024)
Go ahead and axolotl questions