Post-Cutoff.com
  1. Home
  2. Videos
  3. I Tested Sonnet 5.5 vs Opus 5.5 (WILD RESULTS)

I Tested Sonnet 5.5 vs Opus 5.5 (WILD RESULTS)

Brock Mesarich | AI for Non Techies · 2026-09-28 · review · 48,104 views

▶ Watch on YouTube

What's in the video

Description written by Gemini, which watched and listened to the whole video.

Summary
An independent presenter evaluates and benchmarks Anthropic’s Claude Sonnet 5, Sonnet 5.5, Opus 5.5, and Fable 5.1 by having each model generate a full 3D interactive browser game from an identical detailed prompt. He tests the playable outputs in real-time, assessing gameplay, visual quality, and stability while tracking the total generation time and API cost for each model.

What is shown

Claims & numbers

Notable quotes

Assessment
This is an authentic third-party benchmark and hands-on comparison demonstrating the execution of code generated by different LLMs. The generation processes took place prior to recording, but the presenter plays the unedited resulting web games directly in Chrome and displays exact recorded generation durations and API costs.

Described by gemini-3.8-flash on 2026-09-29 from the video's audio and frames.

Related events