mkhazraee/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
A high-throughput and memory-efficient inference and serving engine for LLMs
SGLang is a fast serving framework for large language models and vision language models.
Enhancement Proposals and Architecture Decisions
NVIDIA Inference Xfer Library (NIXL)
100Gbps Intrusion Detection and Prevention System
Advanced Interface Bus (AIB) die-to-die hardware open source
cocotb, a coroutine based cosimulation library for writing VHDL and Verilog testbenches in Python