32:44Claude Opus 5 is a freak
AI Search runs Claude Opus 5 through eight hands on tests, five of them long autonomous runs inside Claude Code, and reports the token cost of every one. It one shots a browser native Windows 11 with a working Excel formula engine, wins an image to 3D render no other frontier model had managed, finds four companies' Q4 filings by itself and turns them into a narrated presentation video, builds an animated X-wing in Blender over MCP, and composes a five minute techno track in a real DAW after choosing and downloading its own synth plugins. It also fails the hidden frog test, gets six out of six brain scans wrong, and underwhelms on deep research. The benchmark tour is where the review turns: first on some independent boards by a single point, third on LiveBench, the slowest frontier model on the table, roughly twice the price of a rival he calls just as intelligent, and a headline ARC-AGI-3 score with a contamination question attached. The verdict is a conditional no, and it is aimed as much at Anthropic as at the model.































































