Reasoning tokens are output tokens. 9 models in the catalogue — every Gemini Flash and Flash-Lite generation, plus Gemini 2.5 Pro — publish a distinct reasoning rate, and all 9 of them set it equal to the model’s own output rate. So thinking bills at the output rate everywhere, and the light bar below is the only part of the answer you ever read. The 48 models with published per-task volumes are listed here; the rest have no token breakdown to split.
ModelAnswer vs. thinking · per taskThinkingOutput $ on thinking
gpt-oss 20BOpenAI · open weights16K89%$0.0029at $0.180 / M out
Nemotron 3 NanoNVIDIA · open weights24K89%$0.0048at $0.200 / M out
Gemini 3.1 Flash-LiteGoogle · reasoning priced separately10K82%$0.0148at $1.50 / M out
GLM-5.2Z.ai · open weights51K79%$0.223at $4.40 / M out
Claude Sonnet 5Anthropic88K75%$0.884at $10.00 / M out
DeepSeek V4 FlashDeepSeek · open weights41K74%$0.0114at $0.280 / M out
Claude Opus 4.8Anthropic52K73%$1.30at $25.00 / M out
DeepSeek V4 Flash 0731DeepSeek · open weights45K73%$0.0600at $1.32 / M out
Qwen3.8 27BAlibaba · open weights48K71%$0.143at $3.00 / M out
Claude Opus 5.5Anthropic84K70%$1.68at $20.00 / M out
K2 Horizon 375B A23BMBZUAI · open weights36K69%—no output rate published
DeepSeek V4 Pro 0813DeepSeek · open weights38K69%$0.151at $3.96 / M out
Muse GlimmerMeta · open weights10K69%$0.0143at $1.50 / M out
GLM-5.3Z.ai · open weights49K69%$0.214at $4.40 / M out
Qwen3.8 2.4T A95BAlibaba · open weights47K68%$0.280at $6.00 / M out
GLM-5.3 FlashZ.ai · open weights47K68%$0.0233at $0.500 / M out
InklingThinking Machines · open weights23K68%$0.0948at $4.05 / M out
Claude Fable 5Anthropic45K67%$2.25at $50.00 / M out
Kimi K3Moonshot AI · open weights32K67%$0.487at $15.00 / M out
Qwen3.8 Flash NextAlibaba · open weights72K67%$0.0337at $0.470 / M out
Gemini 3.5 FlashGoogle · reasoning priced separately36K66%$0.327at $9.00 / M out
DeepSeek V4 Flash VisionDeepSeek46K66%$0.0606at $1.32 / M out
Granite 4.2 8BIBM · open weights22K65%$0.0055at $0.250 / M out
GPT-5.5OpenAI15K65%$0.463at $30.00 / M out
Gemini 2.5 ProGoogle · reasoning priced separately7K64%$0.0677at $10.00 / M out
Kimi K2.7 CodeMoonshot AI · open weights19K64%$0.0767at $4.00 / M out
GPT-6 AstraOpenAI17K61%$0.835at $50.00 / M out
Claude Fable 5.1Anthropic47K60%$2.36at $50.00 / M out
Granite 4.2 3BIBM · open weights11K60%$0.0013at $0.120 / M out
Gemini 3.5 Flash-LiteGoogle · reasoning priced separately10K59%$0.0260at $2.50 / M out
Claude Opus 5Anthropic43K59%$1.07at $25.00 / M out
GPT-5.6 SolOpenAI17K59%$0.346at $20.00 / M out
Nemotron 3 SuperNVIDIA · open weights67K59%$0.0604at $0.900 / M out
MiMo-V2.6-ProXiaomi · open weights38K58%$0.0326at $0.870 / M out
Qwen3.6 27BAlibaba · open weights19K58%$0.0678at $3.60 / M out
gpt-oss 120BOpenAI · open weights15K56%$0.0090at $0.595 / M out
Grok 4.6SpaceXAI19K52%$0.113at $6.00 / M out
Gemini 3.6 FlashGoogle · reasoning priced separately20K48%$0.0752at $3.75 / M out
Muse Spark 1.2Meta23K47%$0.0963at $4.25 / M out
Qwen3.5 397B-A17BAlibaba · open weights9K47%$0.0315at $3.60 / M out
MiniMax M2.7MiniMax · open weights10K46%$0.0117at $1.20 / M out
Nemotron 3.5 LightningNVIDIA · open weights17K46%$0.0037at $0.220 / M out
MiniMax M3MiniMax · open weights22K45%$0.0262at $1.20 / M out
Qwen3.5 122B-A10BAlibaba · open weights7K42%$0.0233at $3.20 / M out
Claude Haiku 4.5Anthropic8K41%$0.0383at $5.00 / M out
Gemini 3.7 FlashGoogle · reasoning priced separately21K36%$0.0796at $3.75 / M out
Qwen3.6 35B-A3BAlibaba · open weights13K36%$0.0281at $2.25 / M out
Qwen3 Coder NextAlibaba · open weights00%$0.0000at $1.20 / M out