Vikrantpalle/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
A high-throughput and memory-efficient inference and serving engine for LLMs
Minimalist ML framework for Rust
Blazing-fast LLM inference in pure Rust. No PyTorch and Python runtime.
MirAIe-AC API for Python
Jambo - JSON Schema to Pydantic Converter
wasmCloud is an open source Cloud Native Computing Foundation (CNCF) project that enables teams to build, manage, and scale polyglot apps across any cloud, K8s, or edge.
wasmCloud Application Deployment Manager (wadm) is a Wasm-native orchestrator for managing and scaling declarative wasmCloud applications.
Generation of diagrams like flowcharts or sequence diagrams from text in a similar manner as markdown
Parser for hls playlist files (*.m3u8)