How Anthropic made Claude 3x faster
Theo - t3․gg · 2026-10-04 · review · 124,529 views
What's in the video
Description written by Gemini, which watched and listened to the whole video.
Summary
Web developer Theo Browne (t3.gg) reviews Anthropic’s engineering blog post, "How we made claude.ai 3x faster in two weeks," which details how Anthropic used autonomous Claude agents in Slack to optimize its web and desktop applications. Theo evaluates the techniques used—such as deterministic metric ratcheting, headless browser profiling, and horizontal sub-agent workflows—while testing Claude.ai live, demonstrating remaining UI cache bugs, and comparing the findings to his own performance engineering on T3 Code and Lakebed.
What is shown
- Anthropic Blog Post Breakdown [02:39]: Reviewing the benchmark table detailing ~3x speed improvements across web, desktop, and Claude Cowork.
- Claude.ai Cache Invalidation Bug [05:43 – 09:20]: Theo demonstrates that deleting threads or refreshing right after sending prompts can cause threads to disappear or fail to update in Claude.ai's sidebar due to over-aggressive IndexedDB caching without proper cache invalidation.
- T3 Code Performance Tooling [10:00 – 10:55, 20:06 – 20:33]: Showcasing hover prefetching and blurred loading states in T3 Chat as alternative patterns for handling cached data without jarring layout shifts.
- Agent Estimation Failure Demo [17:03 – 17:45]: Theo prompts Claude Opus 5.5 in T3 Code to plan a repository migration; the model estimates 25 working days (5 weeks) for an overhaul Theo expects to finish in hours.
- CI Performance Gates [29:10 – 29:55]: Displaying T3 Code’s GitHub Actions automated PR comment table that enforces ceilings on WebSocket/wire transfer payload sizes.
- Layout Shift & Pre-rendering Analysis [38:19 – 38:52, 53:40 – 54:00]: Reviewing Anthropic’s recorded traces of sidebar layout shifts (CLS) and Chrome speculative pre-rendering interactions.
- Table Streaming Comparison [60:28 – 61:38]: Comparing Claude.ai's streaming table rendering cell-by-cell against T3 Code’s approach of buffering content until the table block is fully parsed.
Claims & numbers
- Anthropic’s Reported Speedups: Anthropic claimed making the core user experience of Claude.ai and the desktop app ~3x faster in a two-week sprint during August 2026, merging over 3,000 changes without a customer-facing incident or rollback [03:20, 14:24].
- Core Latencies: P75 time to a typeable page on fresh load reportedly dropped from 3.1s to 0.55s (5.6x faster); desktop app cold start dropped from 6.31s to 3.328s (1.9x faster) [03:52, 04:54].
- Message Sending Speedups: Claude Cowork web message sending latency dropped from 928 ms to 48 ms (19x faster), and Claude Code desktop message latency dropped from 250 ms to 52 ms (4.8x faster) [05:07].
- Target Milestones: Anthropic hit 12 of their 13 initial performance targets by day 3 [16:59].
- Bottlenecks Identified by Claude: Claude ran a React hook census finding 6,900 hooks and 900 store subscriptions in the composer typing path [43:14]; identified a single
:root:has()CSS selector adding 24 ms to every DOM change [43:36]; and found non-Latin characters (like em-dashes) forcing UTF-16 string conversion that froze the main thread for ~0.35s during syntax highlighting [47:45 – 48:02]. - Streaming Budget: Target frame budget was set to 8.33 ms to maintain 120 FPS on 120 Hz displays [54:43, 63:31].
Notable quotes
- [32:27] "With Claude, measuring something makes it tractable... As soon as Claude has a number that it can beat, it can start optimizing."
- [35:36] "If you're just blindly measuring the number of React commits or V8 calls, you're going to start turning things off that don't actually affect performance."
- [57:26] "As soon as you start including words like 'ambitious', or 'boil the ocean', or 'push beyond what I requested'... you're pulling the model out of the happy paths and towards the sketchier, scarier ones."
Assessment
This is an independent technical commentary and critique of an official Anthropic engineering report. The presenter conducts unedited live demonstrations of Claude.ai and T3 Code, uncovering real caching and UI synchronization glitches that demonstrate the trade-offs of the aggressive agent-driven optimizations described in the article.
Described by gemini-3.8-flash on 2026-10-06 from the video's audio and frames.