Post-Cutoff.com
  1. Home
  2. Videos
  3. Sonnet 5.5 (Fully Tested): The MOST USEFUL MODEL…

Sonnet 5.5 (Fully Tested): The MOST USEFUL MODEL YET! RIP ASTRA & SOL!

AICodeKing · 2026-09-29 · review · 5,433 views

▶ Watch on YouTube

What's in the video

Description written by Gemini, which watched and listened to the whole video.

Summary
AICodeKing reviews Anthropic's Claude Sonnet 5.5 (released September 28, 2026), testing it via OpenRouter inside the OpenCode coding-agent harness across the eight interactive tasks of KingBench 3. The video evaluates Sonnet 5.5's code generation, 3D Three.js rendering, algorithmic reasoning, and local model training against Claude Opus 5.5 as a reference standard. Sonnet 5.5 scores 71.5 out of 80 (89.38%), placing just behind Opus 5.5 and GLM 5.3.

What is shown

Claims & numbers

Notable quotes

Assessment
This is an independent hands-on benchmark review evaluating Claude Sonnet 5.5 through a live agent harness. The code generations, interactive Three.js models, gameplay, and terminal scripts are demonstrated directly on screen, with fair and transparent critiques of small mechanical and visual imperfections.

Described by gemini-3.8-flash on 2026-09-29 from the video's audio and frames.

Related events