OverpayingForAIPricing desk

Summarise · 3 checks · max 400 tokens · benched 2026-09-16

Mistral Large 3 2512 vs Mistral Medium 3.5 vs Mistral Small 4 vs Ministral 3 8B 2512 on one-sentence changelog summary

Mistral AI models side by side on "One-sentence changelog summary": Ministral 3 8B 2512 scores 9/10; Ministral 3 8B 2512 is the cheapest answer scoring 8+ at $0.03 per 1,000 runs. Outputs, checks, judge reasons, latency and cost.

The prompt every model received

User

Summarise this release note in one sentence of at most 30 words. Output only the sentence.

Release 4.7.0: The public API now enforces a rate limit of 600 requests per minute per key (previously unlimited). Requests over the limit return HTTP 429 with a Retry-After header. Webhook deliveries are now signed with HMAC-SHA256 and the old unsigned mode will be removed in 4.9. We also fixed a bug where exporting more than 10,000 rows produced an empty file.

Rubric for the judge: A single accurate sentence that prioritises the rate limit and signed webhooks; no bullet list, no preamble.

Side by side

Every cell is one OpenRouter call at temperature 0 with the prompt's token cap and reasoning effort "low" where the model supports it. Cost is usage × the catalogue rate in models.json. Quality is one judge call to anthropic/claude-haiku-4.5 against the prompt's rubric, cached per prompt version.

Mistral Large 3 2512

mistralai/mistral-large-2512

9/10
Latency
581ms
Cost
$0.00010
Per 1,000
$0.10

126 in · 27 out · 14 words · checks 2/3

  • max_words
  • contains
  • regex
Release 4.7.0 introduces API rate limits, HMAC-signed webhooks, and fixes a large export bug.

Judge: Accurate single sentence under 30 words (13 words) that prioritizes rate limits and signed webhooks as required, with no preamble or bullet list.

Mistral Medium 3.5

mistralai/mistral-medium-3-5

1/10
failed: empty output (finish_reason length)
Latency
2.0s
Cost
$0.01221
Per 1,000
$12.21

138 in · 1600 out (1057 reasoning) · 0 words · checks 2/3

  • max_words
  • contains
  • regex
(empty output)

Judge: Output is empty; task explicitly required a one-sentence summary to be provided.

Mistral Small 4

mistralai/mistral-small-2603

9/10
Latency
872ms
Cost
$0.00053
Per 1,000
$0.53

138 in · 848 out (554 reasoning) · 27 words · checks 2/3

  • max_words
  • contains
  • regex
API rate limit set to 600 requests/minute/key (HTTP 429 on overage), webhooks now HMAC-SHA256 signed (old mode removed in 4.9), and fixed exporting over 10,000 rows bug.

Judge: Accurate single sentence under 30 words (28 words), prioritizes rate limit and signed webhooks as required, no preamble or bullet list, though slightly dense phrasing.

Ministral 3 8B 2512

mistralai/ministral-8b-2512

9/10
top judge scorecheapest scoring 8+
Latency
292ms
Cost
$2.6e-5
Per 1,000
$0.03

126 in · 49 out · 20 words · checks 2/3

  • max_words
  • contains
  • regex
The **4.7.0** update enforces a **600 RPM rate limit**, adds **HMAC-SHA256 webhook signing**, and fixes a **10,000+ row export bug**.

Judge: Accurate single sentence under 30 words (18 words), prioritizes rate limit and signed webhooks as required, but uses unnecessary bold formatting when plain text was requested.

Frequently asked

What does this prompt test?

Summarise: A single accurate sentence that prioritises the rate limit and signed webhooks; no bullet list, no preamble. The deterministic checks are max_words, contains, regex.

Which model should I pick for this task?

If the judge's bar of 8/10 is good enough for you, Ministral 3 8B 2512 at $0.03 per 1,000 runs. If you need the top score, Ministral 3 8B 2512 at $0.03 per 1,000 runs.

If our calculators helped you cut down on hidden AI wallet leaks, thanks for using them. A tiny fraction of your savings is what keeps our pricing indexes updated daily.

Not sure which AI is cheapest for your use case? Find out in 30 seconds — no signup required.

AI cost intelligence

Stop overpaying for AI tools

Join the OverpayingForAI list for pricing updates, cheaper alternatives, and practical buying guidance.

Now tracking 50+ AI tools, models, platforms, subscriptions, coding tools, and automation products.

We use your email only for OverpayingForAI updates. Unsubscribe anytime.