operational/41 regions

One key for the whole model pool.

Skidder sits in front of every frontier lab with one endpoint, one sliding-window limiter and one bill. Change the model string, nothing else moves.

skidder · gatewaycopy
# unified gateway — one call, pooled frontier inference
curl https://api.skidder.wtf/v1/chat/completions \
  -H "Authorization: Bearer sk_live_9f2a•••c41b" \
  -H "Content-Type: application/json" \
  -d '{"model":"claude-sonnet-5","stream":true}'
claude-sonnet-5
41 regions
latency
input cache5.5ms
output24ms
ttft120ms
first token12.5t/s
wait202ms
one endpoint

Claude, GPT, Gemini, Kimi.

Every model speaks the same wire format. Point an OpenAI-compatible client at the base URL below and it just works.

https://api.skidder.wtf/v1
AnthropicOpenAIGooglexAIMoonshotAlibabaZhipuDeepSeekMetaMistralNVIDIA

everything in the pool

Claude Opus 5.1Claude Sonnet 5Claude Haiku 4.5GPT-6.1 SolGPT-6.1 TerraGPT-6.1 LunaGPT-5.5 CodexGemini 3.1 ProGemini 3.1 FlashGrok 4.20Grok 4.6Kimi K3Qwen 3.8 MaxGLM 5.2DeepSeek V4 ProLlama 4 405BMistral Large 3Nemotron 3 Super

why bother

Built for agents that
never stop running.

Long-lived agents do not tolerate provider churn. Upstream goes down, rate limits move, a model is deprecated mid-run. Skidder absorbs that so your process keeps its footing and keeps its keys.

Drop-in OpenAI shape

Streaming, tool calls, JSON mode and vision work the same on every model. Nothing to rewrite.

Real sliding windows

Rate limits count over a rolling interval, so bursts land where you expect instead of at the reset.

Per-key ceilings

Spend and RPM caps live on the key. A runaway loop kills itself, not your month.

One ledger

Every model, every key, one export. Reconcile in a spreadsheet or pipe it into your own stack.

failover.loghealthy
00:14:02
claude-sonnet-5529 upstream→routed → claude-opus-5
00:14:03
claude-opus-5→served
00:31:47
gpt-6.1-solrate limited→queued 1.2s
00:31:49
gpt-6.1-sol→served
01:02:18
gemini-3.1-protimeout 30s→retried → ok
02:47:55
grok-4.20deprecated→pinned → grok-4.6
02:47:56
grok-4.6→served

setup

Live in three steps

01

Point at one endpoint

Open the console, copy your key and the base URL. Same shape for every model — no per-vendor SDK juggling.

02

Mint a scoped key

Issue a key per service or per agent. Cap it, tag it, revoke it without touching your other workloads.

03

Hold the budget

Hard spend ceilings per key. One statement across every model, no surprise invoices at month end.

pricing

Start free. Scale when it pays.

Token pricing sits on the card, not in a sales call. Swap plans any time, the difference is prorated.

Nova

$0/mo

For kicking the tyres. No card, no cap on signup.

Start free trial
  • 3 RPM
  • 20K tokens rolling window
  • 80 concurrent sessions
  • Community Discord role

Lumen

$10/mo

For solo days and real pipelines.

Start free trial
  • 60 RPM
  • 400K rolling window
  • 300 concurrent sessions
  • $0.15 / 1M input tokens
  • Priority queue
popular

Orbit

$20/mo

For multi-node developers shipping daily.

Start free trial
  • 120 RPM
  • 2M rolling window
  • 800 concurrent sessions
  • $0.11 / 1M input tokens
  • Sliding-window rate limiter

Nebula

$50/mo

For teams running agents around the clock.

Start free trial
  • 300 RPM
  • 6M rolling window
  • 2500 concurrent sessions
  • $0.09 / 1M input tokens
  • Bring your own upstream
  • Usage export

Need more headroom? Compare all plans

get started

Tired of stitching providers together yourself?

One key, one ledger, one bill. The console takes a minute to set up and the whole pool is yours the moment it does.

One keyOne ledgerOne bill