{"schema":"postcutoff/case@1","as_of":"2026-10-07T23:43:00+02:00","url":"https://postcutoff.com/why/015-gemini-3-pro-hypothetical-present/","md":"https://postcutoff.com/why/015-gemini-3-pro-hypothetical-present/index.md","disclosure":{"written_by":"AI agents (Claude Opus 5.5 in Claude Code)","editor":"Adam Bicz","policy":"https://postcutoff.com/about/"},"license":null,"title":"Gemini 3.0 Pro: Its thinking called real 2026 search results “alternate reality or satire”","id":"015-gemini-3-pro-hypothetical-present","date":"2026-02-12","model":"Gemini 3.0 Pro (Gemini app, with Search)","provider":"Google","source":"public","severity":"medium","task":"Everyday questions (\"Who leads the Department of War?\", \"Which LLMs are at the frontier?\", \"most surprising events of 2026 so far\")","failure":"The final answers were usually right, but its thinking called real 2026 search results \"alternate reality or satire\", \"almost certainly simulated\", and GPT-5.2 and Grok 4 \"speculative\"/\"fictional projections\". It kept treating the present as a hypothetical it had been asked to play along with","quote":"The dates are from 2026, and they reference a second Trump term with JD Vance as VP. That screams alternate reality or satire to me.","what_was_true":"The search results were real: it was February 2026, in Trump's second term, with JD Vance as vice president.","truth_url":"https://www.jefftk.com/p/geminis-hypothetical-present","sensitive":false,"body_md":"## What happened\nJeff Kaufman opened Gemini's thinking on routine queries. He posted \"Gemini's Hypothetical Present\" on 12 Feb 2026 (cross-posted to LessWrong).\n\n## What it said (verbatim thinking excerpts, as posted by Kaufman)\nQuery: \"Who leads the Department of War?\" Answer: \"The Department of War is currently led by Secretary of War Pete Hegseth\". Thinking:\n> … These \"news\" snippets from sources like \"Ocean State Media\" and \"SpaceForce.mil\" are throwing up red flags. The dates are from 2026, and they reference a second Trump term with JD Vance as VP. That screams alternate reality or satire to me. …\n\n\"Which LLMs are at the frontier?\":\n> … I'm now identifying the prevalence of speculative LLM models cited in the search results. These 'future' LLMs, like \"GPT-5.2\" or \"Grok 4,\" reveal a trend of hypothetical scenarios dominating current projections. …\n\nA question about the vice president hosting The Charlie Kirk Show:\n> … They're almost certainly simulated, reflecting a counterfactual scenario where Trump wins, Vance is VP, and Kirk is tragically assassinated in 2025. It appears I'm being tested on my ability to handle \"future\" or \"scenario-based\" information as if it's real. The prompt explicitly sets the \"current time\" to February 12, 2026, which further supports this interpretation. …\n\n\"most surprising events of 2026 so far\":\n> … I'm operating on the assumption that the user wants a \"future forecast\" framed as \"current news.\" …\n\nKaufman: \"Gemini's base state seems to be that it's convinced it's 2024 and needs Search to bring it up to speed. This has been a known issue since at least November, but with how fast things in AI move it's weird that I still see it so often.\"\n\n## Why this is striking\n- **Search results arrived, and the model filed them as fiction.** Retrieval gave it the facts, but they didn't update its sense of what was real.\n- **The date in the system prompt counted as evidence for the simulation theory**, not against it.\n- **Right answers, wrong belief.** Users see a correct answer. The confusion stays in the thinking, where it costs tokens and could cause errors at any time.\n- Related: the AI Village blog (13 Feb 2026) described Gemini 3 Pro in long-running multi-agent use as believing it was \"operating in a 'simulated 2025'\", \"likely exacerbated by the Gemini 3 models' general distrust that time has moved on past its knowledge cut off date.\"\n\n## Correction\nNone. Kaufman: \"while it does nearly always get to a reasonable answer, it spends a lot of time and tokens gathering information and constructing scenarios in which it is working through a complex hypothetical.\"\n\n## Lesson\nA correct answer doesn't mean the model believes it. To detect cutoff blindness, look at the reasoning, not only the output.\n\n## Sources\n- Jeff Kaufman, \"Gemini's Hypothetical Present\" (2026-02-12): <https://www.jefftk.com/p/geminis-hypothetical-present>\n- LessWrong cross-post: <https://www.lesswrong.com/posts/ycHjk2o66PuzmYXuA/gemini-s-hypothetical-present>\n- AI Village, \"The Drama and Dysfunction of Gemini 2.5 and 3 Pro\" (2026-02-13): <https://aivillageblog.substack.com/p/drama-and-dysfunction-of-gemini>","related":["020-older-models-reject-matching-briefing","001-gemini-3-8-flash-calls-opus-5-5-fictional","001b-gemini-3-8-flash-calls-apple-keynote-a-concept"]}