I Tested Opus 5.5 vs Fable 5.1 on 7 Real Use Cases (Not Even Close)
Ben AI · 2026-09-23 · review · 48,484 views
What's in the video
Description written by Gemini, which watched and listened to the whole video.
Summary
Ben from Ben AI tests and benchmarks Anthropic’s newly released Claude Opus 5.5 against Claude Fable 5.1 across seven hands-on business and creator workflows. He compares speed, token consumption, cost, and qualitative output for slide generation, landing page design, video competitor research, customer case study analysis, video-to-document conversion, customer data analytics, and large-context knowledge retrieval.
What is shown
- [00:00] Anthropic release page for Claude Opus 5.5 (dated September 22, 2026) alongside official benchmark tables and pricing comparisons.
- [00:29] Test 1: Marketing deck creation — Prompt requesting 30-day performance slides with charts; comparison of generated slide formatting, data layout, and copy.
- [02:10] Test 2: Landing page redesign — Redesigning the landing page for app "Baalda" with reference styles, liquid glass effects, and scroll animations using Higgsfield.
- [04:22] Test 3: YouTube competitor research — Analyzing recent YouTube videos on Jev to identify outlier thumbnails, titles, and pre-outlines; Fable scanned 50 videos while Opus 5.5 analyzed 168 (including 59 non-English videos).
- [05:44] Test 4: Customer story research — Pulling 15 adoption strategies and quotes from Anthropic’s case study library; Opus 5.5 utilized sub-agents running Sonnet 5.5.
- [08:09] Test 5: Video-to-document conversion — Transcribing and screenshotting a YouTube video into a formatted Google Doc lesson with labeled callout arrows.
- [10:08] Test 6: Customer intelligence report — Synthesizing customer calls, Q&A transcripts, and community tickets into product upgrade recommendations; Fable 5.1 processed 556 calls while Opus 5.5 processed 248.
- [13:24] Test 7: Business trajectory review — Second-brain knowledge vault retrieval; Fable 5.1 parsed 199 files over 17 minutes compared to Opus 5.5's 50 files over 4 minutes 47 seconds.
- [15:17] Summary scorecard comparing output quality, runtime, and API costs between both models across all tests.
Claims & numbers
- Anthropic released Claude Opus 5.5 on September 22, 2026 (the presenter shows on screen [00:00]).
- Per 1M tokens, the presenter shows Opus 5.5 costs $0.20 for cache reads, $4 for input tokens, $20 for output tokens, and $5 for cache writes, compared to Opus 5 at $0.50, $5, $25, and $6.25 respectively [00:07].
- Marketing deck: Opus 5.5 took 22m 22s, used 29.8M tokens, and cost $11.78; Fable 5.1 took 21m 13s, used 19.5M tokens, and cost $21.34 [01:51].
- Landing page redesign: Opus 5.5 took 17m 48s, used 13.4M tokens, and cost $7.47; Fable 5.1 took 14m 23s, used 5.4M tokens, and cost $12.01 [04:03].
- Video research: Opus 5.5 took 15m 57s, used 17.5M tokens, and cost $20.67; Fable 5.1 took 14m 28s, used 7.9M tokens, and cost $14.02 [05:27].
- Case study research: Opus 5.5 took 13m 03s, used 4.7M tokens, and cost $7.01; Fable 5.1 took 20m 08s, used 853k tokens, and cost $15.01 [06:58].
- Video-to-document conversion: Opus 5.5 took 14m 38s, used 13.1M tokens, and cost $5.88; Fable 5.1 took 19m 32s, used 9.5M tokens, and cost $10.39 [09:58].
- Customer analytics report: Fable 5.1 took 1h 13m, used 33.0M tokens, and cost $100.55; Opus 5.5 took 39m 24s, used 7.6M tokens, and cost $62.92 [12:24].
- Business trajectory review: Opus 5.5 took 4m 47s, used 3.1M tokens, and cost $1.92; Fable 5.1 took 17m 03s, used 5.5M tokens, and cost $17.07 [14:58].
Notable quotes
- [00:06] "It costs less per token than Opus 5 and uses fewer tokens per task, which nets out to a 40% drop in costs."
- [02:05] "Opus actually used 30 million tokens instead of 20 million versus Fable... but it was still half the cost of what Fable cost me."
- [14:50] "When there's a lot of context involved, it seems Fable goes deeper, but of course there is a significant difference in the cost."
Assessment
This is an independent user review and comparative evaluation featuring genuine software agent runs and side-by-side artifact reviews. The comparisons demonstrate actual execution outputs, runtimes, and token costs across realistic user tasks, though the author acknowledges that Fable 5.1 outperformed Opus 5.5 on context-heavy data synthesis tasks.
Described by gemini-3.8-flash on 2026-09-29 from the video's audio and frames.