KenForever1/attachment_set
A type safe dynamic attachment container implemented in C++17.
πͺ It's what you do right now that makes a difference.
A type safe dynamic attachment container implemented in C++17.
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++
ζθ°’ζ―ζεε ³ζ³¨οΌKenForever1ηBlog!
llm about docs
vLLM Kunlun (vllm-kunlun) is a community-maintained hardware plugin designed to seamlessly run vLLM on the Kunlun XPU.
Build your own high performance LLM inference engine in C++ and CUDA - a smaller version of vLLM
High-performance, light-weight C++ LLM and VLM Inference Software for Physical AI
A polyphonic harmonica simulator developed based on Rust, supporting real-time keyboard playing, chord playing, tone shifting functions, and built-in automatic demonstration of tracks.
A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.
a cdn image bed
RTSP Stream Tool is an RTSP video streaming tool that provides an intuitive web interface.
Terminal image viewer
a linux memory tool inspired from fincore and vmtouch.
a neovim plugin base on nvim-oxi, rotate the selected character according to the direction and number.
With the increase in the idioms a programmer understands, the language becomes friendlier to them.
Rust bindings to the Triton Inference Server
a thread local var statistic lib implemented by rust
llm_build_from_scratch
A high-performance LLM inference API and Chat UI that integrates DeepSeek R1's CoT reasoning traces with Anthropic Claude models.
LLM inference in C/C++
Python bindings for llama.cpp
A service request benchmark tool used to test the QPS and delay of services.
c sdk export example
LMDeploy is a toolkit for compressing, deploying, and serving LLMs.