Mistral AI 称 Mistral Large 4 在 Harvey 法律智能体基准测试中位居开源模型首位 MMistral AI13小时49分钟前 Mistral AI 称,Mistral Large 4 在 Harvey 的法律智能体基准测试(Legal Agent Benchmark,LAB)中位列开源模型第一。 #mistral-ai#mistral-large-4#基准测试
Mistral AI 公布 Mistral Large 4 在 AutomationBench 上的 657 项工作流测试结果Mistral AI 公布了 AutomationBench 评测结果。在模拟办公应用的 657 项业务工作流中,Mistral Large 4 完成了测试,表现领先于 Kimi K3、MiMo-V2.6-Pro 和 DeepSeek V4 Pro。
Mistral Large 4 在网络安全 CTF 竞速赛中攻克 19 道挑战中的 18 道Mistral AI 表示,Mistral Large 4 在基于实际运行环境的夺旗赛(CTF)竞速测试中通过工具调用攻克了 19 道挑战中的 18 道。
Mistral AI 推出拥有 1 万亿参数的 Mistral Large 4 并立即开放 APIMistral AI 发布了原生多模态模型 Mistral Large 4,总参数量达 1 万亿,激活参数为 490 亿。该模型现已通过 API 及 Mistral Cloud 基础设施提供,并计划于 10 月底开放模型权重。
Mistral AI 演示 Mistral Large 4 在 12 分钟内逆向工程恶意软件Mistral AI 演示了 Mistral Large 4 端到端分析未知二进制文件。该模型在 12 分钟内将样本识别为 Cobalt Strike,提取了失陷指标(IoC)与恶意软件配置,并生成了包含 YARA 规则的报告。
Mistral Large 4 位列 Code Arena: WebDev 榜单第 45 名Mistral Large 4 以 1,534 分在 Arena.ai 的 Code Arena: WebDev 榜单上位列第 45 名,与 Claude Opus 4.8 (High) 仅差 2 分,混合 Token 成本仅为后者的近六分之一。