Mistral Large 4 (ML4), nicknamed Le Chonk, is a natively multimodal model with a 1-trillion-parameter latent and 49 billion active parameters, released as a public preview with weights scheduled at month’s end. Trained from scratch on 3,800 NVIDIA Grace Blackwell GPUs in Mistral’s European datacenters, it’s positioned as an open-weight, sovereign AI that can run on private cloud or on-premise. The preview is being red-teamed with cybersecurity partners and state authorities; the model emphasizes enterprise control, multilingual coverage (training data spanning 160+ languages including all EU official languages), and availability across multiple regions with a fully European deployment option.
Substantively, ML4 is presented as state-of-the-art among open models across cybersecurity, coding, agentic workflows, multimodal vision, science/math, and knowledge work. Key metrics: on cybersecurity it ranks among the global top five on the Artificial Analysis Cyber Index, scores 82% on a vulnerability reproduce-and-patch test, and solves 93% of Cybench challenges. Coding scores include DeepSWE 61.7%, Terminal-Bench 28.3%, and a Coding Agent Index of 49.8%; a blind human coding eval rated it 3.74 (second of five). Agentic and knowledge-work benchmarks include AutomationBench 59.9% and AA-Briefcase 1,393 Elo. Visual grounding exceeds some closed models (Dense200 42% vs 41% for GPT-6-Astra). Scientific and math workflows perform strongly (SciCode-Verified), and safety/robustness show high resistance to prompt-injection attacks (Lakera B3 93.3%, KORA 1.691).
Summary generated by AI from the linked article. hn.today is not affiliated with Hacker News or Y Combinator.