OverpayingForAIPricing desk

Summarise · 3 checks · max 400 tokens · benched 2026-09-16

Qwen3.8 Max (0902) vs Qwen3.8 27B vs Qwen3.8 Flash vs Qwen3.7 Plus on meeting notes to exactly three bullets

Alibaba models side by side on "Meeting notes to exactly three bullets": Qwen3.8 Flash scores 10/10; Qwen3.8 Flash is the cheapest answer scoring 8+ at $0.20 per 1,000 runs. Outputs, checks, judge reasons, latency and cost.

The prompt every model received

System

You are a concise assistant. Output only what is asked for.

User

Turn these meeting notes into exactly three bullet points. Each bullet starts with "- " and is at most 20 words. Output nothing except the three bullets.

Notes: Launch of the new billing page moves from 18 Sept to 25 Sept because the tax calculation for EU customers still fails for VAT-exempt businesses. Priya owns the fix and will demo it Thursday. Marketing will hold the announcement email until the fix is verified. Separately, the team agreed to drop support for the legacy CSV export in Q1 next year and to notify affected customers in November.

Rubric for the judge: Three bullets that capture the delayed launch, the VAT fix owner, and the CSV export decision; nothing extra.

Side by side

Every cell is one OpenRouter call at temperature 0 with the prompt's token cap and reasoning effort "low" where the model supports it. Cost is usage × the catalogue rate in models.json. Quality is one judge call to anthropic/claude-haiku-4.5 against the prompt's rubric, cached per prompt version.

Qwen3.8 Max (0902)

qwen/qwen3.8-max-0902

1/10
Latency
1.6s
Cost
$0.00274
Per 1,000
$2.74

171 in · 400 out (392 reasoning) · 5 words · checks 1/3

  • regex
  • max_words
  • contains
- Billing page launch delayed

Judge: Output contains only one bullet point instead of three required bullets, and fails to capture the VAT fix owner and CSV export decision specified in the rubric.

Qwen3.8 27B

qwen/qwen3.8-27b

1/10
Latency
649ms
Cost
$0.00106
Per 1,000
$1.06

171 in · 400 out (292 reasoning) · 2 words · checks 1/3

  • regex
  • max_words
  • contains


- Launch

Judge: Output contains only one incomplete bullet point instead of three complete, substantive bullets capturing the required information.

Qwen3.8 Flash

qwen/qwen3.8-flash

10/10
top judge scorecheapest scoring 8+
Latency
781ms
Cost
$0.00020
Per 1,000
$0.20

183 in · 362 out (303 reasoning) · 42 words · checks 3/3

  • regex
  • max_words
  • contains
- Billing page launch moves to 25 Sept due EU VAT-exempt tax calculation failures.
- Priya owns the fix, demos Thursday; marketing holds announcement email until verified.
- Team agreed to drop legacy CSV export in Q1 next year, notify customers November.

Judge: Three bullets, each under 20 words, capturing all three required points (delayed launch, VAT fix owner/demo, CSV export decision) with no extra commentary.

Qwen3.7 Plus

qwen/qwen3.7-plus

1/10
failed: empty output (finish_reason length)
Latency
2.6s
Cost
$0.00209
Per 1,000
$2.09

145 in · 1600 out (1600 reasoning) · 0 words · checks 1/3

  • regex
  • max_words
  • contains
(empty output)

Judge: Output is completely empty; task explicitly required exactly three bullet points with specific content.

Frequently asked

What does this prompt test?

Summarise: Three bullets that capture the delayed launch, the VAT fix owner, and the CSV export decision; nothing extra. The deterministic checks are regex, max_words, contains.

Which model should I pick for this task?

If the judge's bar of 8/10 is good enough for you, Qwen3.8 Flash at $0.20 per 1,000 runs. If you need the top score, Qwen3.8 Flash at $0.20 per 1,000 runs.

If our calculators helped you cut down on hidden AI wallet leaks, thanks for using them. A tiny fraction of your savings is what keeps our pricing indexes updated daily.

Not sure which AI is cheapest for your use case? Find out in 30 seconds — no signup required.

AI cost intelligence

Stop overpaying for AI tools

Join the OverpayingForAI list for pricing updates, cheaper alternatives, and practical buying guidance.

Now tracking 50+ AI tools, models, platforms, subscriptions, coding tools, and automation products.

We use your email only for OverpayingForAI updates. Unsubscribe anytime.