Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. Modular pretraining (GRAM) lets dangerous capabilities be…

Modular pretraining (GRAM) lets dangerous capabilities be switched off per module (AE Studio and Anthropic)

★★after cutoffresearchAnthropicAE Studioconfidence: high

On July 8, 2026 researchers from AE Studio and Anthropic published "Modular Pretraining Enables Access Control". It introduces Gradient Routed Auxiliary Modules (GRAM), which route dual-use knowledge such as virology, cybersecurity and nuclear physics into separate modules during pretraining. The modules can be switched on or off, so one training run matches several data-filtered models.

Key facts

What happened

An extension of gradient routing. Dangerous knowledge is placed in removable modules during pretraining instead of being filtered out of the data.

Why it matters

It points to tiered access, where vetted users get a model with, say, the virology module and the public does not, without training separate models. This bears on how labs might ship open or differentiated weights safely.

Changelog

  • 2026-10-01: created (leads run, from the Anthropic uncited-posts audit)

Sources (1)

id: 2026-07-08-anthropic-ae-studio-modular-pretraining-gram · updated 2026-10-01 · open in the interactive timeline