LLM Router live

●
Monitor →
– req/min connecting…

Routing at a glance · waiting for data

Waiting for the first metrics update.

Router

–
port –

Requests today

–
– req/min now

Success rate today

–
– errors

Cloud spend today

–
budget –

Latency today — end-to-end response time

–
average end-to-end

Status codes today

Throughput (req/min, last ~10 min)

Local vs Cloud today

Local – Cloud –
– local · – cloud calls

Token throughput today

–
tokens per second

Fallback depth today

–
won on the first hop

Cost today — call history vs budget

Runtime & feed health

Dashboard feed pipeline–

Sources today — what actually served

Captcha — tile-pick lane and paid solvers

Cooldowns — hops resting after 429 / 5xx

Capacity — local GPU nodes

GPU detail — RTX nodes (main + fallback)

Latency by pool / task / caller / node / modelwhich provider is actually fast
Per-task breakdown today–
Docker containers — MAIN onlyloading…

Cloud share today — which pool served

–

Free usage left — local daily guardrails

–

Provider keys — upstream account health

–
Cloud pools — this run, plus wins todayfree fallback chain
Ready / unused means configured, with no observed attempt this run. Wins confirm a response. Daily caps below are local guardrails; vendor token and rate limits may apply sooner.
#PoolModel Attempts (run)Wins (run) Wins (today) Win share (run)Success (run) 429/4xx5xxLocal daily capServing tasksStatus
Cloud calls per hourtoday · restart-safe

Callers today — who is using the router

CallerRequestsLocal / CloudSuccessTasks
Failures today — non-2xx calls–
Cloudflare Pages · loading traffic scope · data via Cloudflare Worker every ~15s · – · –