QuentinFuxa/WhisperLiveKit
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
https://quentinfuxa.github.io/
Real-time, local speech-to-text with streaming ASR, speaker diarization, translation, and OpenAI/Deepgram-compatible APIs.
Standalone causal streaming runtime for Qwen3-ASR
Simultaneous translation model for 200 languages
Agentic RAG platform purpose-built for small language models (SLM). Robust PDF/SQL search
Research code for AlignAtt4LLM simultaneous speech translation.
Personal portfolio — applied AI engineering, open-source work, and research
France unplugged Qwen for bias. An independent, preregistered audit, as a single-page article.
Nemo prunned for SortFormer only
OCRmyPDF + FastAPI demo: track and display OCR progress in real-time via a custom plugin.
OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
TopicGPT allows to integrate the benefits of LLMs into Topic Modelling
Recursive analysis of python stack of packages using AWS Strands Agents
NVIDIA Canary ASR model optimized for Apple Silicon using MLX.
Community maintained hardware plugin for vLLM on Apple Silicon
Interactive D3.js network graph component for Streamlit — force-directed layout, zone clustering, multiple shapes, search, actions, i18n
LLM inference in C/C++
Fast inference engine for Transformer models
An open-source, code-first Python toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.
Python utility to convert PyTorch model weights from '.bin' to '.safetensors' format.
Add a way to det
model_metric_analysis