joelulu/Awesome-Parallel-Speculative-Decoding
Awesome-Parallel-Speculative-Decoding
Awesome-Parallel-Speculative-Decoding
A high-throughput and memory-efficient inference and serving engine for LLMs
A unified library for building, evaluating, and storing speculative decoding algorithms for LLM inference in vLLM
Awesome-Efficient-Diffusion
[EMNLP 2026] 📚 A curated list of Awesome Efficient dLLMs Papers with Codes
Elevate your AI research writing, no more tedious polishing ✨
Collection of Acceleration Methods for Generative AI
Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.
Official Implementation of LaViDa: :A Large Diffusion Language Model for Multimodal Understanding
The official code that integrates MagCache (Fast Video Generation with Magnitude-Aware Cache) with ComfyUI.
The official GitHub repo for the survey paper "A Survey on Diffusion Language Models".
HunyuanVideo-1.5: A leading lightweight video generation model
Repo no longer maintained
🤗A PyTorch-native Inference Engine with Hybrid Cache Acceleration and Parallelism for DiTs: Z-Image, FLUX2, Qwen-Image, etc.
🌐 Permanent Hosting Site: http://ai-paper-finder.info/ 🌐 Hugging Face Hosting: https://huggingface.co/spaces/wenhanacademia/ai-paper-finder
Collect articles in the GeoAI
Consistency Distillation with Target Timestep Selection and Decoupled Guidance
Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model
HomePage of Huanlin Gao
📚A curated list of Awesome Diffusion Inference Papers with Codes: Sampling, Cache, Quantization, Parallelism, etc.🎉
📚 Collection of awesome generation acceleration resources copy
Repo for the paper "Extrapolating from a Single Image to a Thousand Classes using Distillation"
Awesome Dataset Distillation Papers