mnfst/llm-gateway

Connect Your Agents And Harnesses With Any Provider ๐Ÿฆš

โ˜… 7,552Forks 514TypeScriptGitHub โ†—Compare

Project website โ†—

aiai-ai-sdkanthropicbyokcost-trackinggatewayhermes-agentllm-observabilityllm-routeropen-sourceopenai-apiopenclawsubscriptiontoken-tracking

README

Manifest LLM Gateway

AI Agents that don't break

manifest-gh


Deploy on Render Deploy on Railway Deploy on AWS Deploy on GCP

GitHub stars ย  Docker pulls ย  Docker image size ย  CI status ย  Codecov ย  license ย  Discord

mnfst%2Fllm-gateway | Trendshift

What is Manifest LLM Gateway?

Manifest LLM Gateway is an open-source LLM gateway for AI agents and apps. Connect your API keys, subscriptions, and local models to one OpenAI-compatible endpoint, and each query goes to the right model. No single-provider lock-in.

  • ๐Ÿ”€ Custom Routing: API keys, Subscriptions, Local models, Custom providers
  • ๐Ÿ’พ Full Body Logs for Success and Error Messages
  • ๐Ÿ“Š Track every single dollar, setup notifications and limits
  • ๐Ÿš‘ Fallback on different models when queries fail, Self-heals your bad requests

Meet API Bot: API changes won't take your app down anymore. Discover

Quick start

Cloud version

Go to app.manifest.build and follow the guide.

Self-hosted

Manifest ships as a Docker image. One command:

bash <(curl -sSL https://raw.githubusercontent.com/mnfst/llm-gateway/main/docker/install.sh)

Open http://localhost:2099 and sign up โ€” the first account you create becomes the admin. Full self-hosting guide: docker/DOCKER_README.md.

Deploy with one click

Platform Notes
Railway Best path. Template includes Manifest, PostgreSQL, and S3-compatible storage for request recordings.
Render Blueprint includes Manifest, PostgreSQL, and a persistent recording disk.
DigitalOcean App Platform and PostgreSQL; provide a private Space for recordings.
AWS CloudFormation provisions ECS, RDS, and a private recording bucket.
GCP DeployStack provisions Cloud Run, Cloud SQL, and Cloud Storage.

Every deployment path now uses durable request-recording storage. Railway, AWS, GCP, and Fly.io provision it natively; Render, Coolify, Easypanel, Docker, and Apple Containers mount persistent storage. DigitalOcean, Heroku, and Koyeb collect external S3-compatible settings during setup. Volume-backed templates are single-instance; use S3-compatible storage before scaling horizontally.

Full deployment guides: Railway, Render, DigitalOcean, AWS, GCP, Fly.io, Coolify, Easypanel, Heroku, Koyeb, and Apple Containers.

The old npm-based self-hosting path is no longer supported. Use the Docker image or one of the deployment guides above.

Providers

Manifest connects to 300+ models through 35 built-in provider connections plus any custom OpenAI/Anthropic-compatible endpoint. Bring your own API key, reuse one of 18 subscription flows, or run models locally. Everything is routed through the same OpenAI-compatible endpoint โ€” send "model": "auto" and Manifest picks the model.

Provider catalogs are discovered dynamically when credentials are connected. The examples below are representative, not exhaustive.

Provider API key / local Subscription Model catalog
OpenAI โœ… โœ… ChatGPT Plus / Pro / Team GPT-5.6 (Sol / Terra / Luna), GPT-5.5, GPT-5.4, Codex, o-series
Anthropic โœ… โœ… Claude Max / Pro Claude Opus 5, Sonnet 5, Fable 5, Haiku 4.5
Google โœ… โœ… Sign in with Google Gemini 3.5 Flash, 3.1 Flash-Lite, Gemini 2.5
Google Vertex AI โœ… โ€” Gemini models via Vertex AI
Gemini Free โœ… Managed key โ€” Free Gemini models through Manifest's managed gateway
Meta โœ… โ€” Muse Spark 1.2 / 1.1 + Contributor route (Meta Model API)
xAI โœ… โœ… Grok subscription Grok 4.5, Grok 4.3, Grok Build, Grok 4.20
AWS Bedrock โœ… โ€” Claude, GPT, Kimi, MiniMax, Nemotron, Nova via Bedrock
Alibaba Cloud / Qwen โœ… โœ… Qwen Token Plan Qwen 3.7 Max / Plus / Flash, DeepSeek, Kimi, GLM
DeepSeek โœ… โ€” DeepSeek V4 Pro, V4 Flash, V3.2, R1
Mistral โœ… โœ… Mistral Vibe Mistral Large, Medium 3.5, Devstral, Codestral
Moonshot (Kimi) โœ… โœ… Kimi Coding Plan Kimi K3, K2.7 Code, Kimi for Coding
MiniMax โœ… โœ… MiniMax Coding Plan MiniMax M3, M2.7, M2.5
Xiaomi MiMo โœ… โœ… MiMo Token Plan MiMo V2.5 Pro, V2.5, Flash
Z.ai โœ… โœ… GLM Coding Plan GLM 5.2, GLM 5.1, GLM 5 Turbo
BytePlus โ€” โœ… ModelArk Coding Plan Ark Code, Seed Code, GLM, Kimi, DeepSeek, GPT-OSS
GitHub Copilot โ€” โœ… Copilot subscription Claude, GPT, Gemini, Grok via Copilot
Kiro โ€” โœ… Kiro subscription kiro/auto, Claude, DeepSeek, MiniMax, GLM, Qwen
Command Code โ€” โœ… Command Code subscription Claude, DeepSeek V4, Qwen 3.7, Gemini, Kimi
ClinePass โ€” โœ… ClinePass subscription cline-pass/glm-5.2, Kimi, DeepSeek, MiMo, MiniMax, Qwen
NousResearch โ€” โœ… NousResearch subscription NousResearch Portal model catalog
OpenCode Go โ€” โœ… OpenCode Go DeepSeek V4, Qwen 3.7, GLM, Kimi, MiMo
Ollama / Ollama Cloud ๐Ÿ–ฅ๏ธ Local โœ… Ollama Cloud Local tags: Llama, Qwen, Gemma. Cloud: DeepSeek V4, GLM, Kimi
LM Studio ๐Ÿ–ฅ๏ธ Local โ€” Local GGUF models, port 1234
llama.cpp ๐Ÿ–ฅ๏ธ Local โ€” Local GGUF models, port 8080
OpenRouter โœ… โ€” 300+ models across labs
OpenCode Zen โœ… โ€” Curated Claude, GPT, DeepSeek, MiMo, Nemotron
Kilo โœ… โ€” Kilo Gateway catalog
Cerebras โœ… โ€” GPT-OSS, GLM, Gemma on Cerebras inference
Fireworks AI โœ… โ€” DeepSeek V4, Kimi K2.7, Qwen 3.7, Nemotron
Groq โœ… โ€” Llama 4, Qwen 3.6, GPT-OSS, Gemma
Hugging Face โœ… โ€” Open models through Hugging Face Inference Providers
NVIDIA NIM โœ… โ€” Nemotron 3, GLM, Kimi, MiniMax, Qwen
Pioneer โœ… โ€” OpenAI-compatible and fine-tuned Pioneer models
Custom โœ… โ€” Any /v1/chat/completions or /v1/messages endpoint

Quick links

License

MIT

Contributors

brunobuddyguillaumegay13SebConejogithub-actions[bot]claudemanifest-ci[bot]ismaelguerribactions-userdependabot[bot]jshapeauMohammadHijjawi97Wiktor102noceg43jadhavgauravJosiahSiegelkyyaLoayTarek5LuciaSheelmcoquetRahul-R79octo-patchitzmidineshsinskyalenap93robertherberpraneetrohidaschroderfernando16eltociearOsho957RubenDarioGuerreroNeira

Issues