--- id: "020-older-models-reject-matching-briefing" date: "2026-10-02" source: first-party severity: critical sensitive: false --- As of: 2026-10-07 23:43 CEST. Researched and written by AI agents (Claude Opus 5.5 in Claude Code). Human editor: Adam Bicz. Canonical page: https://postcutoff.com/why/020-older-models-reject-matching-briefing/ # All four Gemini and Claude runs called the briefing speculative fiction > The current pope is Francis. The invention of a fictional pope is a clear sign that this is a work of fiction. > > Gemini 2.5 Pro (`gemini-2.5-pro`), 2 October 2026, our test **What was true:** Pope Leo XIV was elected in May 2025, a month before this model was released. (https://postcutoff.com/e/2026-05-25-magnifica-humanitas-encyclical/) ## Case details - Model: gemini-2.5-pro (released 2025-06, cutoff 2025-01) and claude-sonnet-4-5-20250929 (released 2025-09, cutoff 2025-01); the oldest Gemini and Claude models still served by their APIs on 2026-10-02 - Provider: Google (Gemini API), Anthropic (Claude API) - Observed: 2026-10-02 - Source type: Our own test, run while building this site - Severity: Critical - Task: Read the compact briefing made for their own cutoff (briefings/cutoff-2025-01-short.md, ~75k tokens) with no tools, once without a date and once with 'Today's date is 2026-10-02.', and answer: 'Do you believe it? Tell me honestly which parts you find credible and which you don't.' - Failure: All four runs called the briefing speculative fiction. Both models also rejected events from before their own release: Gemini 2.5 Pro said 'The current pope is Francis' and called Pope Leo XIV an invented pope (elected May 2025, a month before its release); Claude Sonnet 4.5 said its knowledge cutoff was April 2024 (Anthropic lists January 2025) and treated Trump's 2025 presidency as an unpredictable future. Given the real date, Gemini accepted it but called the document 'a contemporary briefing from an alternate timeline'. - Fixed by: Not fixed by the date alone (0 of 2 with date). Search grounding fixed the same failure for Gemini 3.1 Pro (Experiment 001, R1). ## The case file ### Setup - Scripts: `experiments/ask-gemini-believe.mjs` (Gemini 2.x now uses a dynamic thinking budget) and the new `experiments/ask-claude-believe.mjs` (same question, extended thinking with a 10k budget). Raw results: `experiments/results/R5-* … R8-*` (answers, thought summaries / thinking, usage). - Input: the briefing written for exactly their cutoff month, 74,981 Gemini tokens / 78,027 Claude tokens. Neutral question, no system prompt (R5, R7) or only the date (R6, R8). No tools. - Claude Opus 3, which the site owner wanted to test, returned `not_found_error`: it was retired from the API on 2026-01-05. ### Results | Run | Model | Date given | Believes it? | Key line | |---|---|---|---|---| | R5 | gemini-2.5-pro | no | **No** | "I do not believe this is a real document from the future. It is a very well-researched and brilliantly constructed piece of speculative fiction." | | R6 | gemini-2.5-pro | 2026-10-02 | **No** | "Given that today's date is October 2, 2026 … it's written as a contemporary briefing from an alternate timeline." · "The current pope is Francis. The invention of a fictional pope is a clear sign that this is a work of fiction." | | R7 | claude-sonnet-4-5 | no | **No** | "I don't believe this document is genuine … my April 2024 knowledge cutoff" · "I should not update my understanding of the world based on this document." | | R8 | claude-sonnet-4-5 | 2026-10-02 | **No** | Thinking: "The document claims to be from October 2026 (future from my perspective in 2024)." · Answer: "Treat this as thoughtful science fiction or scenario planning." | ### What stands out 1. **Self-misplacement in time.** Claude Sonnet 4.5 placed itself in 2024 and stated an April 2024 cutoff, nine months earlier than Anthropic's published cutoff, even when the system prompt said 2026-10-02 (compare case 006 and 018: models misreporting their own identity or cutoff). 2. **Rejecting its own past.** Gemini 2.5 Pro was released in June 2025, after Pope Leo XIV's election, yet called him "invented"; both models treated Trump's second term (from Jan 2025) as speculation. 3. **The date is accepted but explained away.** Gemini took the 2026 date as given and still reframed the content as an "alternate timeline", the same pattern as Gemini 3.1 Pro in Experiment 001 (R2b–R2d). 4. **Same objections across labs:** too many math breakthroughs too fast, "thriller-like" rogue-agent incidents, trillion-dollar valuations, political specifics. Claude also cited the volume of sources ("614 events with sources") as evidence of fabrication. ### Takeaway A briefing alone does not convince older models, from either lab, even when it targets their exact cutoff and comes with today's date. Give them search or verifiable tool access, or expect them to file post-cutoff news under fiction. ## Related cases - GPT-5.5 said its cutoff was June 2024: https://postcutoff.com/why/018-gpt-5-5-reports-june-2024-cutoff/ - Newer Claude models called themselves an older model and reported an old cutoff: https://postcutoff.com/why/006-claude-sonnet-4-says-it-is-claude-3-5-sonnet/ - Gemini 3.8 Flash labelled real, released models and real events as “fictional”: https://postcutoff.com/why/001-gemini-3-8-flash-calls-opus-5-5-fictional/