Management Consulting Tasks (Internal)Browse 296

Management Consulting Tasks (Internal)

Can the model produce useful work in a professional knowledge-work setting?

Version not specifiedExact reported variant
Knowledge Work7 ranked models7 reported values1 reportsHigher is better

Top rankings

One row per model, using its best reported score across effort settings.

7
RankModelBest scoreBest reported setting
1GPT-5.6 SolOpenAI43.2%Reported configurationOpenAI report1 value · 1 reportJul 9, 2026 · source
2GPT-5.6 TerraOpenAI37.2%Reported configurationOpenAI report1 value · 1 reportJul 9, 2026 · source
3Claude Fable 5Anthropic35.5%Reported configurationOpenAI report1 value · 1 reportJul 9, 2026 · source
4GPT-5.6 LunaOpenAI35.4%Reported configurationOpenAI report1 value · 1 reportJul 9, 2026 · source
5Claude Opus 4.8Anthropic31.6%Reported configurationOpenAI report1 value · 1 reportJul 9, 2026 · source
6GPT-5.5OpenAI31.3%Reported configurationOpenAI report1 value · 1 reportJul 9, 2026 · source
7Gemini 3.1 Pro PreviewGoogle13.2%Reported configurationOpenAI report1 value · 1 reportJul 9, 2026 · source

Effort curve

Every sourced cost-linked effort value for this exact version. Lines connect complete sweeps only.

0
No cost-linked effort sweep for this version.
Definition and comparison boundaryinternal methodology

Can the model produce useful work in a professional knowledge-work setting?

Professional analysis, document, finance, office, or domain-specific tasks. Higher is better. The value is the percentage reported in this lab's table.

This is a publisher-defined internal evaluation. The task set or grading details are not fully public, so treat it as directional evidence.

Method / source