~ same prompt, different LLM — every model gets the same brief in Blender, you judge the render ~
OpenAI / GPT 6 Astra / thinking:max has entered the arena NEW!
Render Window — hero
hero render
camera + framing set up by the AI contestant
3D Viewport — param-tower User Persp
LOADING 3D VIEW
$0.04Cost
2245sDuration
46Tool calls
47Turns
Blender 5.1.0MCP 1.0.02026-07-31
Tokens: 962736/43239Vision: Votes W-L-T: 6-0-0reasoning: effort maxDONE ✓
B-BENCH — TRANSCRIPT OpenAI / GPT 5.6 Luna / thinking:max · r1 · 2026-07-31
RUN LOG — param-tower · OpenAI / GPT 5.6 Luna / thinking:max · r1 · ok
SYSTEM PROMPTthe control — every model on this task got these instructions
PROMPTblender-bench v8HARNESSraw-v3

WITHHELD · The prompt text is unavailable: the publish security scan withheld this run's transcript, which carried it. The prompt versions above still identify exactly which prompt this run was given.

>
FULL CALL INSPECTOR
OPEN RAW JSONL ↗
0 SHOWN
LOADING COMPLETE TRANSCRIPT…
Votes cast on this bench: 000482 · 157 runs · 25 contestants
Best viewed with Netscape Navigator 4.07 at 800×600 · real runs, live ratings