~ same prompt, different LLM — every model gets the same brief in Blender, you judge the render ~
OpenAI / GPT 6 Astra / thinking:max has entered the arena NEW!
Render Window — hero
hero render
camera + framing set up by the AI contestant
3D Viewport — ravensthwaite-terminal User Persp
LOADING 3D VIEW
$2.97Cost
1747sDuration
Tool calls
Turns
Blender 5.1.0MCP 2026-09-06
Tokens: 1381/59095Vision: Votes W-L-T: 0-0-0reasoning: effort maxNO DONE
B-BENCH — TRANSCRIPT OpenAI / GPT 6 Astra / thinking:max · r2 · 2026-09-06
RUN LOG — ravensthwaite-terminal · OpenAI / GPT 6 Astra / thinking:max · r2 · ok
SYSTEM PROMPTthe control — every model on this task got these instructions
PROMPTblender-bench v8HARNESSoneshot-v3

WITHHELD · The prompt text is unavailable: the publish security scan withheld this run's transcript, which carried it. The prompt versions above still identify exactly which prompt this run was given.

FULL CALL INSPECTOR
OPEN RAW JSONL ↗
0 SHOWN
LOADING COMPLETE TRANSCRIPT…
Votes cast on this bench: 000482 · 157 runs · 25 contestants
Best viewed with Netscape Navigator 4.07 at 800×600 · real runs, live ratings