Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. Anthropic Frontier Red Team: frontier models reach…

Anthropic Frontier Red Team: frontier models reach superhuman photo geolocation and can write working drone strike software

★★★after cutoffpolicy-safetyAnthropicMoonshot AIconfidence: high

On Sept 10, 2026 Anthropic's Frontier Red Team published new evaluations of AI in tactical intelligence targeting (geolocating people from photos and posts, linking accounts) and conventional weapons (simulated drone terminal guidance, payload drops, GPS-denied navigation). Mythos-class models beat the best human GeoGuessr baseline at photo geolocation, and Opus 5 wrote drone code that struck a moving vehicle on 47% of launches. The open-weights Kimi K3 trailed the frontier but showed "concerning" capability. Anthropic has added classifiers that block weapons-development requests.

Key facts

What happened

The report came out the same day as Anthropic's September threat intelligence report, which covers misuse of Claude for conventional weapons and surveillance. The Frontier Red Team built capability evaluations along the "kill chain" (find, fix, track, target, engage, assess). The intelligence evaluations used real data with hidden ground truth. The weapons evaluations had models write guidance, navigation and control code for simulated drones under wind, clutter, camouflage and GPS jamming or spoofing. Opus 5 did better than the Mythos-class models on several drone tasks. Anthropic attributed this to engineering habits, such as more careful tracking.

Why it matters

It is one of the first public, quantitative assessments by a frontier lab of LLM uplift for surveillance and weapons engineering. Before this, published misuse evaluations had focused on cyber and bio. It supports Anthropic's argument, repeated in its later GLM-5.3 cyber report, that open-weights models without safeguards spread dangerous capabilities. All results come from simulations and Anthropic's own evaluations.

Changelog

  • 2026-09-30: created (Anthropic blog audit; the post had not been cited)

Related events

  1. Anthropic threat intelligence report: AI-orchestrated cyberattacks and distillation by Chinese labs ★★★
  2. Anthropic Frontier Red Team: open-weights GLM-5.3 nearly matches Claude Mythos Preview at exploit development, with weak safeguards ★★★★
  3. Moonshot AI releases Kimi K3, a 2.8T-parameter open-weights multimodal model ★★★★★
  4. Anthropic releases Claude Opus 5 — near-Fable-5 intelligence at half the price ★★★★
  5. Anthropic releases Claude Fable 5 and Claude Mythos 5 — first generally available Mythos-class model ★★★★★
  6. US and Russia strip human review of AI-selected targets from the draft UN autonomous-weapons text ★★★

Sources (3)

id: 2026-09-10-anthropic-intelligence-targeting-weapons-evals · updated 2026-09-30 · open in the interactive timeline