TonyLianLong/LLM-groundedVideoDiffusion
[ICLR 2024] LLM-grounded Video Diffusion Models (LVD): official implementation for the LVD paper
Post-training Researcher at Thinking Machines Lab | UC Berkeley 26' (EECS PhD) | UC Berkeley 22' (CS)
[ICLR 2024] LLM-grounded Video Diffusion Models (LVD): official implementation for the LVD paper
LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models (LLM-grounded Diffusion: LMD, TMLR 2024)
AnimeGAN.js: Photo Animation for Everyone
A gradio web UI demo for Stable Diffusion XL 1.0, with refiner and MultiGPU support
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.
Remodeled LaTeX-style personal website
Official Implementation of the CrossMAE paper: Rethinking Patch Dependence for Masked Autoencoders
Swin-Transformer-based version of RIDE (ICLR 2021 Spotlight)
Games for tomorrow's programmers.
An ARM emulator ported to WebAssembly and running Linux in browsers.
[CVPR 2023] Segmenting objects in videos without human annotations 🤯: Official implementation for Bootstrapping Objectness from Videos by Relaxed Common Fate and Visual Grouping
spi flash driver in linux user space using gpio.
Implementation for VPBench proposed in paper Visually Prompted Benchmarks Are Surprisingly Fragile
Use Hugging Face with JavaScript
Improved Implementation for Training GLIGEN: Open-Set Grounded Text-to-Image Generation
The web-based visual programming editor.
[ECCV 2022] Official Implementation for Unsupervised Selective Labeling for More Effective Semi-Supervised Learning
Democratizing Reinforcement Learning for LLMs
Scaling RL on advanced reasoning models
Nano vLLM
An iOS App which can let you run Windows or Linux on your iOS devices without jailbreak, based on ported version of bochs.
[CVPR 2021] Official Implementation of VAI: Unsupervised Visual Attention and Invariance for Reinforcement Learning
SGLang is a fast serving framework for large language models and vision language models.
😎 An up-to-date & curated list of awesome semi-supervised learning papers, methods & resources.
My learning notes/codes for ML SYS.
Utils for Unsloth
🤗 Diffusers: State-of-the-art diffusion models for image and audio generation in PyTorch
一个萌萌的小创意——给你的脸增加猫耳朵(纯web前端, Vue.js)
Website for Trevor Darrell's group at Berkeley