Whamp/pi-extensions
Will Hampson's Pi coding agent extensions: pstack, quiet, and personal packages
Motorcycles. Dogs. Bonds. Autodidact. http://course.fast.ai alum
Will Hampson's Pi coding agent extensions: pstack, quiet, and personal packages
Pi extension for async subagent delegation with truncation, artifacts, and session sharing
Pi extension for CliffCompaction mechanical context summaries
Public agent skills authored and maintained by Will Hampson.
Make Pi sessions feel endless
CliffCompaction — autocompaction for long-horizon coding agents
Enable/disable skills from loading into pi context at startup
Web-based client for https://herdr.dev/ terminal session manager
GVS5H: Five Qwen3.8-27B Models Match Claude Fable 5 on LiveCodeBench Hard | Fable 5 Level Coding for a Fifth the Price - or on a Single GPU
RLM (Recursive Language Model) extension for pi - process large context files that exceed LLM context windows
Pi harness for DeepSWE benchmark experiments
A high-throughput and memory-efficient inference and serving engine for LLMs
🧻 Knowledge system for agents. Local-first, file-based, progressively disclosed.
Beautiful, Modern & Opinionated Linux
Community recipes for serving LLMs on RTX 3090. Multi-engine (vLLM, llama.cpp, SGLang) and model-agnostic. Currently shipping Qwen3.6-27B configs for 1× and 2× cards.
Ampere DeepSeek V4 FlashMLA fork
Benchmarking Large Language Models using the Eleusis card game
A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support and full compatibility with vLLM, SGLang, and Transformers.
A safetensors extension to efficiently store sparse quantized tensors on disk
Transformers-compatible library for applying various compression algorithms to LLMs for optimized deployment with vLLM
Annotate and review coding agent plans and code diffs visually, share with your team, send feedback to agents with one click.
A simple tool for coordinating several AI agents.