Thursday, 06 August 2026 · World
USD/EUR 0.8657 USD/GBP 0.7425 USD/JPY 157.6 USD/CNY 6.762 All rates →
RSS
EUROS The World Financial Report
Nº 26 Thursday, 06 August 2026 · World Edition
LATEST
Deals & M&A

Qwen 3.8-Max and Claude Opus 5 show why raw benchmark scores don't predict the bill

VentureBeat · 2h ago
Qwen 3.8-Max and Claude Opus 5 show why raw benchmark scores don't predict the bill

Alibaba released Qwen 3.8-Max this week and marketed the preview as second only to Claude Fable 5 (their launch-day table was more equivocal: the model leads on one of 12 coding-agent rows). But an independent harness came close to the opposite conclusion: a benchmark run , apparently using the Preview version, put Qwen 3.8-Max's best effort setting mid-pack, and its default setting last. Both results are real and defensible. The gap between them is about token and time budgets, and that matters because those figures aren’t usually headline numbers. Alibaba's footnotes give its codin

Read the full story at VentureBeat →