Benchmark · as of 18 August 2026

AI Cost Benchmark 2026: What a Completed Task Costs

38 current models, 12 business scenarios, vendor list prices. The tables show the machine share of a completed task in euros and, where the calculator knows a solve rate, the cost per solved task.

  • 38

    current models

  • 12

    business scenarios

  • 2

    scenarios with documented solve rate

The range per scenario

Between the cheapest and the most expensive current model lies, for every one of the twelve scenarios, more than a factor of one hundred. The median shows where the field stands; the factor shows how much model choice matters.

The range per scenario
ScenarioModelscheapest modelMedianmost expensive modelFactor
Answering a support ticket38Ministral 3 3B 0.001 €0.008 €gpt-5.4-pro 0.23 €462×
Reading an email and drafting the reply38Ministral 3 3B 0.000 €0.005 €gpt-5.4-pro 0.14 €479×
Writing or revising a text38Ministral 3 3B 0.001 €0.015 €gpt-5.4-pro 0.51 €636×
Sales email and proposal draft38Ministral 3 3B 0.001 €0.023 €gpt-5.4-pro 0.75 €576×
Classification and routing38Ministral 3 3B 0.000 €0.003 €gpt-5.4-pro 0.06 €318×
Extracting data from documents38Ministral 3 3B 0.002 €0.034 €gpt-5.4-pro 0.84 €352×
Knowledge-base question38Ministral 3 3B 0.001 €0.017 €gpt-5.4-pro 0.47 €394×
Research agent over documents38Ministral 3 3B 0.016 €0.282 €gpt-5.4-pro 13.10 €814×
Writing or changing code38Ministral 3 3B 0.005 €0.087 €gpt-5.4-pro 3.27 €605×
Working through a long document25gpt-5.4-nano 0.097 €0.858 €gpt-5.4-pro 28.47 €295×
Meeting notes38Ministral 3 3B 0.002 €0.029 €gpt-5.4-pro 0.75 €375×
Translation38Ministral 3 3B 0.001 €0.013 €gpt-5.4-pro 0.53 €1,059×

The cheapest models per scenario

For each scenario the five cheapest models by machine cost, plus the cheapest model in each class. The input sizes are listed below each table; they are example sizes from the calculator presets, not measurements.

Answering a support ticket

Input 3,700 tokens, output 400 tokens, one call per task, 2 minutes of rework

Answering a support ticket
ModelClassMachine per task
Ministral 3 3BEdge class0.0005 €
Ministral 3 8BEdge class0.0008 €
Qwen3.5 9B (Together)Compact class0.0009 €
Mistral Small 4Compact class0.0010 €
Ministral 3 14BEdge class0.0011 €
Cheapest model per class
Mistral Large 3Frontier class0.0032 €
Qwen3.7-Plus (Together)Workhorse class0.0022 €
Qwen3.5 9B (Together)Compact class0.0009 €

Reading an email and drafting the reply

Input 2,000 tokens, output 300 tokens, one call per task, one minute of rework

Reading an email and drafting the reply
ModelClassMachine per task
Ministral 3 3BEdge class0.0003 €
Ministral 3 8BEdge class0.0004 €
Qwen3.5 9B (Together)Compact class0.0005 €
Mistral Small 4Compact class0.0006 €
Ministral 3 14BEdge class0.0006 €
Cheapest model per class
Mistral Large 3Frontier class0.0019 €
Qwen3.7-Plus (Together)Workhorse class0.0013 €
Qwen3.5 9B (Together)Compact class0.0005 €

Writing or revising a text

Input 4,000 tokens, output 1,200 tokens, one call per task, 5 minutes of rework

Writing or revising a text
ModelClassMachine per task
Ministral 3 3BEdge class0.0008 €
Ministral 3 8BEdge class0.0012 €
Qwen3.5 9B (Together)Compact class0.0015 €
Ministral 3 14BEdge class0.0016 €
Mistral Small 4Compact class0.0020 €
Cheapest model per class
Mistral Large 3Frontier class0.0058 €
Qwen3.7-Plus (Together)Workhorse class0.0044 €
Qwen3.5 9B (Together)Compact class0.0015 €

Sales email and proposal draft

Input 6,000 tokens, output 1,200 tokens, one call per task, 20 minutes of rework

Sales email and proposal draft
ModelClassMachine per task
Ministral 3 3BEdge class0.0013 €
Ministral 3 8BEdge class0.0019 €
Ministral 3 14BEdge class0.0026 €
Qwen3.5 9B (Together)Compact class0.0026 €
Mistral Small 4Compact class0.0030 €
Cheapest model per class
Mistral Large 3Frontier class0.0088 €
Qwen3.7-Plus (Together)Workhorse class0.0067 €
Qwen3.5 9B (Together)Compact class0.0026 €

Classification and routing

Input 1,500 tokens, output 30 tokens, one call per task, one minute of rework

Classification and routing
ModelClassMachine per task
Ministral 3 3BEdge class0.0002 €
Mistral Small 4Compact class0.0003 €
Ministral 3 8BEdge class0.0003 €
Qwen3.5 9B (Together)Compact class0.0003 €
gpt-5.4-nanoCompact class0.0004 €
Cheapest model per class
Mistral Large 3Frontier class0.0010 €
Qwen3.7-Plus (Together)Workhorse class0.0007 €
Mistral Small 4Compact class0.0003 €

Extracting data from documents

Input 15,000 tokens, output 600 tokens, one call per task, 5 minutes of rework

Extracting data from documents
ModelClassMachine per task
Ministral 3 3BEdge class0.0024 €
Ministral 3 8BEdge class0.0036 €
Mistral Small 4Compact class0.0041 €
Qwen3.5 9B (Together)Compact class0.0042 €
Ministral 3 14BEdge class0.0049 €
Cheapest model per class
Mistral Large 3Frontier class0.0131 €
Qwen3.7-Plus (Together)Workhorse class0.0087 €
Mistral Small 4Compact class0.0041 €

Knowledge-base question

Input 8,000 tokens, output 400 tokens, one call per task, 3 minutes of rework

Knowledge-base question
ModelClassMachine per task
Ministral 3 3BEdge class0.0012 €
Ministral 3 8BEdge class0.0018 €
Mistral Small 4Compact class0.0020 €
Qwen3.5 9B (Together)Compact class0.0023 €
Ministral 3 14BEdge class0.0024 €
Cheapest model per class
Mistral Large 3Frontier class0.0065 €
Qwen3.7-Plus (Together)Workhorse class0.0048 €
Mistral Small 4Compact class0.0020 €

Research agent over documents

Caution case, see cost per solved task below

Input 25,000 tokens, output 1,200 tokens, 8 calls per task, 5 minutes of rework

Research agent over documents
ModelClassMachine per task
Ministral 3 3BEdge class0.0161 €
Ministral 3 8BEdge class0.0242 €
Mistral Small 4Compact class0.0315 €
Ministral 3 14BEdge class0.0322 €
gpt-5.4-nanoCompact class0.0479 €
Cheapest model per class
Mistral Large 3Frontier class0.0968 €
DeepSeek V4 FlashWorkhorse class0.0968 €
Mistral Small 4Compact class0.0315 €

Writing or changing code

Input 15,000 tokens, output 1,500 tokens, 3 calls per task, 10 minutes of rework

Writing or changing code
ModelClassMachine per task
Ministral 3 3BEdge class0.0054 €
Ministral 3 8BEdge class0.0081 €
Ministral 3 14BEdge class0.0108 €
Mistral Small 4Compact class0.0113 €
Qwen3.5 9B (Together)Compact class0.0136 €
Cheapest model per class
Mistral Large 3Frontier class0.0341 €
DeepSeek V4 FlashWorkhorse class0.0293 €
Mistral Small 4Compact class0.0113 €

Working through a long document

Input 300,000 tokens, output 3,000 tokens, one call per task, 15 minutes of rework

Working through a long document
ModelClassMachine per task
gpt-5.4-nanoCompact class0.0965 €
Gemini 3.1 Flash-LiteCompact class0.1236 €
Gemini 3.5 Flash-LiteCompact class0.1516 €
gpt-5.6-lunaCompact class0.1898 €
DeepSeek V4 FlashWorkhorse class0.2114 €
Cheapest model per class
DeepSeek V4 ProFrontier class0.6342 €
DeepSeek V4 FlashWorkhorse class0.2114 €
gpt-5.4-nanoCompact class0.0965 €

Meeting notes

Input 15,000 tokens, output 800 tokens, one call per task, 3 minutes of rework

Meeting notes
ModelClassMachine per task
Ministral 3 3BEdge class0.0020 €
Ministral 3 8BEdge class0.0031 €
Mistral Small 4Compact class0.0035 €
Qwen3.5 9B (Together)Compact class0.0036 €
Ministral 3 14BEdge class0.0041 €
Cheapest model per class
Mistral Large 3Frontier class0.0113 €
Qwen3.7-Plus (Together)Workhorse class0.0075 €
Mistral Small 4Compact class0.0035 €

Translation

Input 2,000 tokens, output 2,000 tokens, one call per task, 2 minutes of rework

Translation
ModelClassMachine per task
Ministral 3 3BEdge class0.0005 €
Ministral 3 8BEdge class0.0008 €
Ministral 3 14BEdge class0.0010 €
Qwen3.5 9B (Together)Compact class0.0011 €
Mistral Small 4Compact class0.0019 €
Cheapest model per class
Mistral Large 3Frontier class0.0052 €
Qwen3.7-Plus (Together)Workhorse class0.0041 €
Qwen3.5 9B (Together)Compact class0.0011 €

Cost per solved task where a solve rate is documented

Where a continuously maintained benchmark series provides a solve rate per model, the calculator prices in attempts, error cost and review time. For the other scenarios no such series exists, and no number is shown on purpose.

Answering a support ticket

Benchmark series: tau2 · 2 models with a value

Answering a support ticket
ModelSolve rateAttemptsMachine per taskper solved task
gpt-5.273 %1.60.025 €1.07 €
gpt-5.159 %2.00.022 €1.34 €

Research agent over documents

Benchmark series: browsecomp · 15 models with a value

Research agent over documents
ModelSolve rateAttemptsMachine per taskper solved task
gpt-5.6-sol92 %1.31.207 €11.01 €
Claude Opus 591 %1.32.749 €13.95 €
gpt-5.6-terra88 %1.40.506 €15.01 €
Claude Sonnet 585 %1.40.867 €18.17 €
Claude Mythos 588 %1.44.189 €18.19 €
gpt-5.6-luna83 %1.40.053 €18.75 €
DeepSeek V4 Pro83 %1.40.259 €18.86 €
gpt-5.584 %1.41.279 €18.88 €

Method

  1. Prices are vendor list prices in USD per 1 million tokens, without discounts, checked on 18 August 2026 on the vendors' pricing pages. Converted at the ECB reference rate of 1.1576 USD per euro.
  2. Machine cost per task: input, output and cache shares following the logic of the AI cost calculator, one attempt, no batch processing, no router, no US residency. Models whose context window cannot hold the input are omitted from that table.
  3. Cost per solved task: machine cost divided by the solve rate from the named benchmark series, corrected for dependent retries (factor 1.2), plus error and review cost according to the scenario profile. Where no series exists, nothing is interpolated.
  4. Only models with an active, bookable endpoint and a published list price are included.
  5. Your own sizes, cache shares, batch processing and error costs can be set in the calculator.

Changes

  • 18 August 2026: first edition with 38 current models. Updated quarterly with the calculator's price status.

How to cite

Convios GmbH (2026): AI Cost Benchmark 2026, as of 18 August 2026, https://www.convios.com/en/insights/ki-kosten-benchmark-2026. Free to use with attribution (CC BY 4.0).