onthehub97/recipes
Hardware-pinned, verified recipes for serving large open-weight LLMs. Every number has a method.
Hardware-pinned, verified recipes for serving large open-weight LLMs. Every number has a method.
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.