Official models · up to 90% off

The LLM router built for
Claude Code and Codex

https://api.teamorouter.cn/v1
BASE URL
MODEL PROVIDERS
CODING AGENTS & CLIENTS
Claude Code
Codex
CC Switch
OpenClaw
Works with 10+ coding agents and clients
  • Smart Routing
Claude
GPT
Gemini
  • Smart Routing
  • Prompt Caching
  • Auto Fallbacks
  • One Bill
  • Spend Tracking
  • Never Downgraded

Trusted by 6,000+ dev teams and indie builders

Live pricing · up to 90% off

Prices update in real time and move with upstream costs. Each request is billed at the discount in effect when it's made. All prices in USD per 1M tokens.See live discounts

Model Context Input (List) Output (List) Input (TeamoRouter) Output (TeamoRouter) Provider Uptime (SLA)
GPT-5.6 SolOpenAI 1M $5.00Cache $0.50 $30.00 $0.53 90% offCache $0.053 $3.18
Claude Fable 5Anthropic 1M $10.00Cache $1.00 $50.00 $2.42 76% offCache $0.242 $12.10
Claude Sonnet 5Anthropic 1M $3.00Cache $0.30 $15.00 $0.50 -84%Cache $0.050 $2.52
Claude Opus 4.8Anthropic 1M $5.00Cache $0.50 $25.00 $1.33 -74%Cache $0.133 $6.65
GPT-5.5OpenAI 1M $5.00Cache $0.50 $30.00 $0.55 -89%Cache $0.055 $3.30
Gemini 3.1 Pro (Preview)Google 2M $2.00Cache $0.20 $12.00 $0.29 -86%Cache $0.029 $1.73
GPT-5.6 TerraOpenAI 1M $2.50Cache $0.25 $15.00 $0.27 90% offCache $0.027 $1.61
GPT-5.6 LunaOpenAI 1M $1.00Cache $0.10 $6.00 $0.11 90% offCache $0.011 $0.64
GPT-5.4OpenAI 1M $2.50Cache $0.25 $15.00 $0.27 90% offCache $0.027 $1.62
GPT-5.4 miniOpenAI 1M $0.75Cache $0.075 $4.50 $0.09 -89%Cache $0.009 $0.51
DeepSeek V4 ProDeepSeek 1M $0.435Cache $0.0036 $0.87 $0.435Cache $0.0036 $0.87
GLM-5.2GLM 1M $1.40Cache $0.26 $4.40 $1.40Cache $0.26 $4.40
139B+tokens routed yesterday
>99%prompt cache hit rate
99.98%routing SLA
Hoursto new-model support

Pay Less for Every Token

Pay only for what you use. Model rates start at just 10% of the official price — discounts applied automatically, per model.

Credits never expire · Buy credits anytime

Quick start

Point your existing tools at TeamoRouter. A one-line change — official SDKs, standard endpoints, nothing to relearn.

View API docs
main.py

Point the base URL at TeamoRouter, drop in your API key — the rest of your code stays the same.

Running in Production

All three numbers come from live production traffic — the same data behind the pricing table and leaderboard on this page.

PRICE 90%off

6,000+ providers compete on price in real time under continuous quality monitoring, so every request gets the best available rate. Live pricing for every model is published on this page.

TIME TO FIRST TOKEN 2.4s

Latency and throughput on par with going direct to the provider, measured across all production traffic. Slow routes are pulled from rotation within minutes; the live numbers are on this page.

UPTIME 99.98%

Our Harness Routing engine backs a production-grade SLA and cache-hit guarantee: automatic failover across providers, around the clock, no downtime.

A Router You Can Build Your Business On

Traditional gateways weren't built for agent workloads. TeamoRouter is tuned for Claude Code, Codex, and long agentic sessions — without sacrificing the basics.

Model Integrity Get exactly the model you request. Every request is routed to the model and protocol you select through vetted upstream providers, never silently swapped or downgraded.
Data Privacy Prompts and completions are processed only for routing, metering, billing, abuse prevention, and support. Logs are never used to train models or sold as usage data.
Spend Tracking Billing details show every request, including model, token usage, applied discount, and final charge. Teams can see exactly where credits are spent instead of guessing from aggregate spend.

How are the prices this low?

One account, one bill. We vet providers, so you don't have to.

Providers compete for your traffic Model providers compete for every request; only those that pass our quality gates are eligible. You get market rates without doing the price discovery yourself.
Buy at scale We secure enterprise-scale volume commitments with vetted model providers. That is where the discount comes from.
Investor-backed pricing Backed by leading funds, we're passing part of that on as price subsidies — a limited-time boost on top of volume discounts.

Production-grade routing

Routes that pass strict quality validation and latency testing — prioritizing availability, response speed and result consistency.

  • PriceUp to 92% off
  • LatencyFirst-token latency on par with official APIs
  • Best forProduction workloads · critical paths · quality-sensitive work
  • AssuranceStrict validation · continuous live monitoring & replacement

Live discounts

Discounts are repriced every hour as upstream costs move. Each request is billed at the discount in effect when it's made.

No matching models Try another keyword, or tap All to clear filters
NOW
AVG
BEST
NET INPUT
NET OUTPUT
NET CACHE

48 hours · hourly · dashed = Average

Usage leaderboard

Updated daily at 00:00 UTC+8

Daily token volume

Total tokens routed through TeamoRouter per day

354.3Btokens ↗ 11.0% vs yesterday

Top agents

Ranked by tokens routed this week · change vs last week

  • 1. Codex Desktop9817.1K requests 1170.1Btokens↗ 30.2%
  • 2. Claude CLI1793.1K requests 268.2Btokens↗ 75.1%
  • 3. VSCode863.8K requests 107.4Btokens↗ 45.4%
  • 4. Claude Desktop517.5K requests 56.9Btokens↗ 104.3%
  • 5. OpenClaw189.1K requests 19.9Btokens↗ 151.2%

Top models

Ranked by tokens routed this week · change vs last week

  • 1. GPT-5.6 SolTTFT 1908ms · 99.46% success 1068.3Btokens↗ 31.8%
  • 2. GPT-5.6 TerraTTFT 2133ms · 98.56% success 251.6Btokens↗ 99.5%
  • 3. GPT-5.5TTFT 1562ms · 98.70% success 136.8Btokens↘ 19.3%
  • 4. Claude Opus 4.8TTFT 2371ms · 99.35% success 109.1Btokens↗ 122.6%
  • 5. GPT-5.6 LunaTTFT 2074ms · 98.93% success 97.7Btokens↗ 57.6%

Updated daily at 00:00 UTC+8 · Last updated

Cut Your AI Spend up to 90%

Top up in the dashboard and usage is drawn down per request — so the budget you set goes further.

Credits never expire · Buy credits anytime

FAQ

A single API that sits between your tools and model providers. You call one endpoint with one key; the router sends each request to Claude, GPT, Gemini or others, meters usage, and gives you one bill.

Yes — both formats. The OpenAI-compatible endpoint works with official OpenAI SDKs and Codex by changing the base URL. Claude Code connects through the Anthropic-compatible endpoint with two environment variables.

Volume. We commit to enterprise-scale usage with vetted providers and pass the difference through. Discounts vary per model and move with upstream costs — the pricing table above is live.

No. Prompts and completions are processed only for routing, metering, billing, abuse prevention and support. They are never used for training and never sold.

No. Buy credits when you want, use them whenever. No subscription required.

TeamoRouter

Every Request Takes The Best Route

TeamoRouter