Claude Opus 5.5 Might Be The Best!!! (3D, Web Design, Animation)
Codex Community · 2026-09-22 · review · 92,003 views
What's in the video
Description written by Gemini, which watched and listened to the whole video.
Summary
Adrian Twarog reviews Anthropic’s Claude Opus 5.5, evaluating its capabilities in agentic coding, complex web design, 3D development, and automation integrations. He examines community examples before running four separate coding prompts in Claude, inspecting the generated websites, UI animations, and functional dashboard.
What is shown
- [00:02] Benchmark charts comparing Claude Opus 5.5 against Fable 5.1, Opus 5, GPT-6 Astra, and GPT-5.6 Sol across Terminal-Bench 4.0, FrontierCode v1.1, and CursorBench 4.0.
- [00:20] Community showcases on X: Blender 3D procedural scene generation, Unreal Engine underwater game creation via Higgsfield, rigged and animated octopus in Blender, and claymation generation.
- [01:24] Prompt 1: Generating an interactive showcase website teaching users about Claude Opus 5.5 with GSAP/Three.js; inspecting the resulting particle sphere animation, thinking-effort toggles, and token economics display at [01:49].
- [03:00] Prompt 2: Redesigning an existing website (
typeui.sh); inspecting original versus generated redesign featuring interactive sound effects, brand kits, dark/light themes, and UI animations at [03:48]. - [04:57] Prompt 3: Building a 3D space agency website using Three.js; inspecting the interactive rocket assembly wireframe, launch sequence, and planetary flyby animation at [05:25].
- [06:18] Prompt 4: Integrating the Zapier SDK to build a personal daily monitoring dashboard; showing the resulting interface with email summaries, YouTube metrics, and connected API tools at [07:29].
- [08:11] Adrian discussing execution speeds, thinking times (often 45–60 minutes per large generation), and overall design output quality.
Claims & numbers
- The presenter claims Claude Opus 5.5 is 30% faster and 40% cheaper than previous Opus models.
- Benchmark screen claims Claude Opus 5.5 scores 66.4% on Terminal-Bench 4.0 (at maximum effort, listed at $7.35), 54.4% on FrontierCode v1.1, and 57.8% on CursorBench 4.0.
- The presenter states API list prices drop 20% to $4 per million input tokens and $20 per million output tokens, with cache reads dropping 60% to $0.20 per million tokens.
- The presenter notes Opus 5.5 thinking mode cannot be toggled completely off (adaptive thinking by default), and "medium" effort on Opus 5.5 is comparable to "high" effort on Opus 5.
- The presenter claims each complex coding task took around 45 to 60 minutes of model reasoning and execution time (e.g., 48m 11s, 51 minutes).
Notable quotes
- [01:43] "It ran for an hour, which is incredibly long compared to previous examples of it creating websites like this."
- [04:52] "This is essentially what I would expect from a professional graphics designer."
- [08:48] "It's almost like handing it off to a person and waiting for them to come back and give you an answer on whatever they've been tasked to do."
Assessment
This is an independent user review and hands-on capability demonstration of Claude Opus 5.5 using local developer environments and the Claude UI. While generation waiting times are edited down, the output code, interactive front-ends, and 3D scenes are demonstrated live in the browser.
Described by gemini-3.8-flash on 2026-09-29 from the video's audio and frames.