A blazing fast AI Gateway with integrated guardrails. Route to 1,600+ LLMs, 50+ AI Guardrails with 1 fast & friendly API.
-
Updated
May 25, 2026 - TypeScript
A blazing fast AI Gateway with integrated guardrails. Route to 1,600+ LLMs, 50+ AI Guardrails with 1 fast & friendly API.
Fastest enterprise AI gateway (50x faster than LiteLLM) with adaptive load balancer, cluster mode, guardrails, 1000+ models support & <100 µs overhead at 5k RPS.
External-model router for Codex with guided Kimi OAuth/API, DeepSeek, safe migration, and rollback.
Model router for agentic systems. Routes every prompt to the right model in <50ms. Cut costs 40-70% with just an endpoint change.
Open-source desktop AI agent workspace with one-click Claude Code, Codex, OpenClaw, Hermes Agent setup and custom LLM model routing.
One Endpoint, Every Model. Route, Monitor, and Failover — All From Your Terminal.
Analyze how LLM model routers can be hacked and how to secure them.
Universal LLM router for AI coding tools. Works with Claude Code, Cursor, Codex, Gemini CLI, Copilot and more. Free-first fallback chain keeps costs 70–85% lower.
Claude Code hooks that auto-switch model tier based on task complexity
Open-source AI backend control plane for model routing, agent tools, OpenAI-compatible APIs, and private workspaces.
Local OpenCode-backed LLM gateway for OpenAI, Anthropic, Gemini, and Responses API-compatible tools, with streaming and tool/function calling.
Automatically route each Pi prompt to a model based on task complexity, price, speed, or context length. Smart model router for pi.
The Fastest enterprise AI gateway — route LLM traffic across OpenAI, Anthropic, Gemini, Groq & 30+ providers via a single API. Self-hosted, no vendor lock-in. ( 55x faster than litellm )
一个本地 AI 网关,统一管理模型请求 · Local AI gateway for any AI client — route, convert, fail over, and trace requests across 26 providers.
Automated quality, cost, and latency evaluation of Microsoft Foundry Model Router against any baseline model — bring your own prompts, get a full report in one command.
ModelFlux: OpenAI-compatible model traffic router with health-aware key-pool scheduling and failover
模型路由与成本优化器:简单问题 flash 直答、故障自动降级、会话 token/缓存/成本实时面板 | Model router & cost optimizer for DeepSeek Harness: flash quick-answers for simple questions, failure fallback, live token/cache/cost panel
Go LLM gateway — one interface for Claude Code, Codex, Gemini CLI, Anthropic, OpenAI, Qwen, and vLLM.
Proxy patch for Google Antigravity that intercepts internal Cloud Code API and adds support for OpenAI, Claude, Ollama, Gemini and any custom LLM provider.
Add a description, image, and links to the model-router topic page so that developers can more easily learn about it.
To associate your repository with the model-router topic, visit your repo's landing page and select "manage topics."