Ross ROSS = Recommend OSS · open-source software intelligence for agents

diegosouzapw/OmniRoute

Never stop coding. Free MIT AI gateway: one endpoint, 350 providers (90+ free), 1200+ models Kimi, Claude, GPT, Gemini, GLM, DeepSeek, MiniMax. Works with Claude Code, Codex, Cursor, OpenCode, Cline & Copilot. Quota-aware auto-fallback, RTK+Caveman compression saves 15-95% tokens, MCP/A2A, Desktop/PWA. Built by 450+ contributors observed · 2026-08-28

github.com/diegosouzapw/OmniRoute · homepage · TypeScript · MIT (permissive) observed · 2026-08-28

Health v2 · maintenance only

78/100

  • Activity 99
  • Release rhythm 87
  • Longevity 14
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: 0.0
  • age_days: 201
  • days_rel: 7
  • days_push: 7
  • n_releases_24m: 279

Full methodology

Adoption not part of the score

56223 stars · 7731 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

OmniRoute is a free, MIT-licensed AI gateway that exposes a single OpenAI-compatible endpoint routing to 350+ LLM providers (90+ with free tiers) and 1200+ models, with quota-aware auto-fallback and RTK+Caveman token compression. It ships with a web dashboard, Desktop/PWA app, MCP/A2A support, and works with coding agents like Claude Code, Codex, Cursor, Cline, and Copilot.

Use cases

  • route claude code to free llm providers
  • avoid hitting api rate limits with automatic provider fallback
  • save tokens when using ai coding agents
  • use one endpoint for openai anthropic and gemini models
  • access free llm tiers without a credit card
  • self-host an openai-compatible proxy for multiple providers
  • connect cursor or cline to gemini and deepseek models

When to choose

  • you use AI coding agents and want quota-aware failover across many providers
  • you want to stack free LLM tiers behind one OpenAI-compatible endpoint
  • you need API translation between OpenAI, Claude, and Gemini formats
  • you want token compression to cut costs on tool-heavy agent sessions

When to avoid

  • you need a single dedicated provider with guaranteed SLAs and enterprise support
  • you require strict data residency or cannot route prompts through third-party free tiers
  • you only use one local model and don't need routing or fallback

Facets

service · maturity active

api-gateway llm-inference proxy rate-limiting middleware mcp chat-interface developer-tools large-language-models developer-tools self-hosted apis backend cross-platform self-hosted cli llm-gateway ai-gateway openai-proxy auto-fallback token-compression free-tier claude-code cursor openai-compatible multi-provider a2a pwa ai-agents nodejs docker web-server

4 sources

Member repositories

RepositoryRoleHealth v2
diegosouzapw/OmniRoutemain78

For agents

markdown · JSON · MCP: product_card(name="diegosouzapw/OmniRoute")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem