Skip to content
TokenGauge
Menu

Model index · 15 models · October 2026

Every model in the index, by vendor.

Each entry lists the SayGM price, the OpenRouter price with its top-up fee, and the workload the model suits. Open a model for its full spec sheet.

Model catalogue grouped by vendor. Prices in USD per million tokens, snapshot from October 2026.
ModelBest forSayGMin / outBlended80/20vs OpenRouterfee incl.
Anthropic
Claude Fable 5.1FrontierThe hardest reasoning and agent tasks, where cost matters least$8.89 / $44.45$16.00216% lower
Claude Opus 5.5FrontierAgentic coding, long multi-step tasks, hard reasoning$3.56 / $17.78$6.40416% lower
Claude Sonnet 5.5FrontierDaily coding in Cursor, Cline and Claude Code$1.78 / $8.89$3.20216% lower
Claude Opus 5FrontierAgentic coding, long multi-step tasks, hard reasoning$4.45 / $22.23$8.00616% lower
Claude Sonnet 5FrontierDaily coding in Cursor, Cline and Claude Code$1.78 / $8.89$3.20216% lower
Claude Haiku 4.5FrontierClassification, extraction, high-volume pipelines$0.89 / $4.45$1.60216% lower
OpenAI
GPT-6.1 SolFrontierGeneral assistants and agents at mid-tier cost$1.49 / $7.45$2.68229% lower
GPT-6 LunaFrontierChat products, summarisation, cheap tool calls$0.07 / $0.37$0.13033% lower
GPT-5.5FrontierReasoning-heavy tasks, research agents$3.73 / $22.35$7.45429% lower
GPT-5.6 TerraFrontierGeneral assistants, balanced cost and quality$1.49 / $8.94$2.98029% lower
GPT-5.6 LunaFrontierChat products, summarisation, cheap tool calls$0.15 / $0.89$0.29829% lower
Google
Gemini 3.1 Pro PreviewFrontierLong context, multimodal input, document analysis$1.39 / $8.34$2.78034% lower
Gemini 3.5 FlashFrontierFast multimodal tasks, high-throughput apps$1.04 / $6.26$2.08434% lower
Moonshot AI
Kimi K3Open weight · TEE variantOpen-weight coding and agent workloads$0.98 / $4.88$1.76021% lower
DeepSeek
DeepSeek V4.1 FlashOpen weight · TEE variantBulk generation, cheap open-weight inference$0.18 / $0.72$0.288122% higher

Blended price assumes 80% input and 20% output tokens. Open-weight OpenRouter figures use the cheapest listed route.

§ 01

How the prices spread

The same models on one chart. The gap is widest on Gemini and narrowest on Claude, and it reverses on input for two open-weight models.

Fig. 1Blended price per million tokens, 80% input and 20% output
  • SayGM
  • OpenRouter, fee included
Show data as a table
ModelSayGMOpenRouter, fee includedDifference
Claude Fable 5.1$16.00$18.99-16%
GPT-5.5$7.45$10.55-29%
Claude Opus 5$8.01$9.50-16%
Claude Opus 5.5$6.40$7.60-16%
GPT-5.6 Terra$2.98$4.22-29%
Gemini 3.1 Pro Preview$2.78$4.22-34%
Claude Sonnet 5.5$3.20$3.80-16%
Claude Sonnet 5$3.20$3.80-16%
GPT-6.1 Sol$2.68$3.80-29%
Gemini 3.5 Flash$2.08$3.16-34%
Kimi K3$1.76$2.24-21%
Claude Haiku 4.5$1.60$1.90-16%
GPT-5.6 Luna$0.298$0.422-29%
GPT-6 Luna$0.13$0.194-33%
DeepSeek V4.1 Flash$0.288$0.13122%
Log scale, so equal distances mean equal percentage gaps. The right column is the SayGM price relative to OpenRouter. Source: published rate cards, October 2026 snapshot. Current prices on saygm.com (opens in a new tab).

§ 02

Confidential open-weight models

Two models in the index also come as a sealed -tee variant that runs inside an Intel TDX enclave, so prompts are decrypted only there.

  • Kimi K3

    The largest discount in the SayGM catalogue: 67.5 percent below Moonshot's list price on the open tier. A sealed -tee variant runs inside a TEE for prompts that must stay confidential.

  • DeepSeek V4.1 Flash

    40 percent below DeepSeek's list price on SayGM's open tier. OpenRouter's cheapest provider is cheaper still, but not confidential. The sealed -tee variant runs inside an Intel TDX enclave.

What the enclave protects · Compare any two models

§ 03

Beyond this index

The index tracks ten widely used models. SayGM lists about 65, including the full Claude, GPT and Gemini lines and more open-weight options. OpenRouter lists several hundred, so check both catalogues if a niche or legacy model matters.