CodeSOTAIntelligence
Methodology

Assumptions & Methodology

Every number on this site is either measured from data we collect, or assumed with a stated value. This page lists all of them so you can audit, disagree, and re-run with your own inputs.

updated August 05, 2026 · data 2024-11-09 → 2026-08-05 (635 days, 792 models)

1 · Where the data comes from

sourcecovershow
OpenRouter frontend statslast ~10–31 days, exactmodel-activity per model: prompt/completion tokens, requests, tool calls — full per-model granularity. Scraped daily.
OpenRouter author pages~90-day rolling, per modelEach lab page embeds a chart of daily tokens for its top models. Parsed for the recent window.
Internet Archive (Wayback)back to 2024-11-09Historical snapshots of the same author pages, stitched together to extend history ~1.5 years.
OpenRouter app pagessnapshot (daily fwd)Per-app top-20 model breakdown — the demand side. Totals, not a daily series.
Model list / pricingcurrentPer-model prompt & completion list prices, free flag, creation date.

2 · The assumptions (and their values)

assumptionvaluerationale & how it's used
Gross token margin assumed90%Inference cost ≈ 10% of list price. Used for all "gross profit" figures. Linear lever — swap freely.
OpenRouter take-rate assumed5–10%Credit fees + BYOK fee + provider spread on GMV. We show a 5 / 7.5 / 10% range rather than one number.
OR → total-API scale assumed×10OpenRouter is estimated at ~10% of a major lab's API traffic. Used only for "est. total revenue". The biggest single uncertainty — measured "via OR" figures don't use it.
Frontier training cost assumed$100MOrder-of-magnitude compute cost of one Opus-class run. Excludes R&D, staff, serving. Used for model ROI.
Prompt:completion mix (history) assumedrecent actualAuthor-page history gives total tokens only. For historical revenue we apply each model's real recent prompt:completion split (from the exact data) at its actual list price. Recent-window revenue is fully exact.
Useful-life threshold assumed50% of peakA model's "useful life" = launch → first day usage falls below half its peak. Choice of threshold; 50% is conventional.
Smoothing method7-day MACentered 7-day moving average on all time-series to remove weekend seasonality.
Partial-day trim method<40% of medianThe current UTC day is incomplete; trailing days below 40% of the weekly median are dropped so charts don't dip to zero.
Free / preview exclusion methodexcludedFree and preview variants spike then vanish (misleading). Excluded from launch/fade/head-to-head analysis.
China vs West classificationby lab originDeepSeek/Qwen/MiniMax/Moonshot/Z.ai/Xiaomi/Tencent/StepFun/ByteDance = China; OpenAI/Anthropic/Google/xAI/Meta/Mistral/Microsoft = West. Stealth/unknown (OpenRouter's Owl) excluded.

3 · Known limitations

How to read the site: trust the measured columns as fact; treat assumed figures as a transparent frame you can re-dial. Every "est. total" / "profit" / "OpenRouter revenue" number is downstream of the assumptions above.