APEX-Agents: Management Consultant

Created by experts from McKinsey, BCG, Deloitte, Accenture, EY
Muse Spark 1.1
Muse Spark 1.1xHigh
63.1%
Opus 5
Opus 5Max
62.6%
GPT-5.6 Sol
GPT-5.6 SolMax
60.4%
Grok 4.6
Grok 4.6High
59.8%
Kimi K3
Kimi K3Max
59.7%
GPT-5.5
GPT-5.5xHigh
59.4%
Opus 4.8
Opus 4.8Max
57.1%
Fable 5
Fable 5Max
57.1%
GPT-5.2
GPT-5.2xHigh
56.4%
GPT-5.4
GPT-5.4xHigh
55.9%
Opus 4.7
Opus 4.7Max
55.4%
GPT-5.3-Codex
GPT-5.3-CodexHigh
54.8%
Sonnet 5
Sonnet 5High
54.5%
GLM-5.2
GLM-5.2
54.3%
Gemini 3.1 Pro
Gemini 3.1 ProHigh
54.2%
DeepSeek-V4-Flash
DeepSeek-V4-FlashMax
54.1%
Gemini 3 Pro
Gemini 3 ProHigh
52.6%
Opus 4.6
Opus 4.6High
51.8%
Grok 4.5
Grok 4.5High
51.0%
GPT-5.2-Codex
GPT-5.2-CodexHigh
50.7%
Kimi K2.7 Code
Kimi K2.7 CodeAuto
50.1%
Opus 4.5
Opus 4.5High
43.7%
Applied Compute: Small
Applied Compute: Small
37.8%
GLM-4.7
GLM-4.7
21.2%
0%
10%
20%
30%
40%
50%
60%
70%
80%

APEX NEWSLETTER

The latest on frontier AI performance, straight to your inbox.

New benchmarks, leaderboard shifts, and research from the APEX team.

By subscribing you agree to receive updates from Mercor.
Unsubscribe anytime.