# Mistral AI 公布 Mistral Large 4 於 AutomationBench 上的 657 項工作流程測試結果

Mistral AI 公布了 AutomationBench 評測結果。在模擬辦公應用的 657 項業務工作流程中，Mistral Large 4 完成了測試，表現領先於 Kimi K3、MiMo-V2.6-Pro 和 DeepSeek V4 Pro。

Language: zh-HK
Time zone: Asia/Hong_Kong

HTML: https://didcodexreset.com/zh-hk/news/2ac55c89f47a4e05a81a8484.html

Content language: zh-HK
Localization state: translated

來源：Mistral AI · 發布於 7/10/2026 22:33:03

Mistral AI 在 AutomationBench 上評測了 Mistral Large 4，測試了 Gmail、Google Sheets、Slack 和 Salesforce 等模擬辦公應用中的 657 項業務工作流程。

據 Mistral AI 表示，Mistral Large 4 的表現領先於 Kimi K3、MiMo-V2.6-Pro 和 DeepSeek V4 Pro 等競爭模型。

Tags: Mistral AI, Mistral Large 4, 基準測試

[查看主帖原文](https://x.com/MistralAI/status/2107834895674323354)
