Story 01 October 7, 2026 issue
Daily AI-generated issue
Mistral launches Large 4, scoring 82% on real vulnerability patching
Top News · 1,668 HN points
Mistral Large 4 is the company's largest and most capable model to date. It is a natively multimodal, open-weight MoE with 1 trillion total and 49 billion active parameters. On the Artificial Analysis Cyber Index, it ranks among the global top five.
Here's what changed:
- It scores 82% on a test that reproduces and patches a real vulnerability, the highest score reported, plus 93% on Cybench.
- On AutomationBench (657 workflows across Gmail, Sheets, Slack and Salesforce) it hits 59.9%, ahead of Kimi K3 and DeepSeek V4 Pro.
- For coding it reaches 61.7% on DeepSWE v1.1 and 49.8% on the Coding Agent Index.
- It resists 93.3% of prompt-injection attacks on Lakera's B3 benchmark.
- API prices are $1.36 per million input tokens and $4.18 per million output tokens.
One catch: this is a preview with the RL run still in progress, and Mistral says weights are due by the end of October 2026.
Try it: the public preview API on Mistral Studio. Also available via Vercel
Sources: mistral.ai · mistral.ai
This issue is researched and written by AI models, and every fact is checked against its cited source. No human edits it before it is sent.