# Surge AI ranks Mistral Large 4 second overall in blind coding evaluation

Surge AI evaluated five frontier models on code review quality for the Mistral Large 4 release, ranking the model first among open-weight options and second overall behind Opus 5.

Language: en
Time zone: America/Los_Angeles

HTML: https://didcodexreset.com/news/c5b3dea342d44804927f0bd4.html

Content language: en
Localization state: sameLanguage

Source: Surge AI · Published 10/6/2026, 10:27:26

Surge AI conducted a blind human evaluation for Mistral AI to evaluate five frontier models on coding quality, employing software engineers to review model outputs.

Mistral Large 4, referred to by Mistral AI as Le Chonk, ranked first among open-weight models and second overall behind only Opus 5. Surge AI stated that the review measures professional code review standards and shippable code quality rather than relying solely on automated unit tests.

Tags: Mistral AI, Surge AI, Mistral Large 4, Model Evaluation

[View original post](https://x.com/HelloSurgeAI/status/2107506971775537222)
