Claude Sonnet 5
answered
~3.1 s for this call: OpenRouter p50 latency + output time, busiest provider, fetched 2026-09-24
jev
answered
327 ms: its p50 over 45 live calls we timed on 2026-09-24 (TypeSafe reports 70–500 ms)
Lane lengths are to scale for these two figures. Your model and inputs will differ: the shadow run measures them.
Claude Sonnet 5 · per call$0.001000$1,000.00 per million calls
jev · per call$0.0000241$24.11 per million calls
difference97.6% lower41.5× cheaper · 9.6× faster
A 400-input, 20-output-token call at OpenRouter list prices fetched 2026-09-24: Claude Sonnet 5 at $2 in / $10 out per 1M tokens, Jev at $0.042 in, output free. Jev bills its own framing: 574 input tokens for this call, from 259 + 31 per question + 0.71 × input, measured on live calls.
Independent check: a test of 1,000 eight-way routing decisions found Jev about 40× cheaper than GPT-5.6 Terra ($0.0151 vs $0.6089). AY Automate