diegosouzapw/omniroute
Never stop coding. Free MIT AI gateway: one endpoint, 359 providers (150+ free), 1200+ models Kimi, Claude, GPT, Gemini, GLM, DeepSeek, MiniMax. Works with Claude Code, Codex, Cursor, OpenCode, Cline & Copilot. Quota-aware auto-fallback, RTK+Caveman compression saves 15-95% tokens, MCP/A2A, Desktop/PWA. Built by hundreds of contributors
- Stars
- 73.7k
- Gained, 7 days
- +2,223
- Gained, 30 days
- n/a
- Reddit posts
- 3
TypeScript · MIT · last push 2026-10-06 · created 2026-02-13
GitHub stars, last 90 days
Does it work with local models?
Reports from people who ran it with a model on their own hardware. Tool calling depends on the model as much as on the server, so say which model and quant you used.
No reports yet.
Add your report
Sign in to say whether it worked for you.
Reddit posts linking it
- [Showcase] OmniRoute ships an MCP server (95 tools, 30 scopes, 3 transports) that lets agents drive an entire AI gateway — routing, quota, compressionr/mcp · 2026-07-02
- A free gateway that puts Ollama and 237 cloud providers in one fallback ladder — local-first with cloud overflow (MIT, self-hosted)r/ollama · 2026-07-02
- I built a free, self-hosted gateway so I never hit a Claude limit mid-session — it drains your subscription first, then falls back across 237 providers (MIT)r/ClaudeAI · 2026-07-02
In the official MCP registry
No registry entry points at this repository, so how it runs is not listed here. Check its README.