Post-Cutoff.com
  1. Home
  2. Models
  3. MolmoAct 2 / MolmoAct 2-Think

MolmoAct 2 / MolmoAct 2-Think

Ai2 (Allen Institute for AI)currentroboticsMolmoActopen weights

Checkpoints: MolmoAct2 (post-trained multi-embodiment foundation, ~5.4B params per HF safetensors), -Think, -Pretrain, fine-tuned -DROID, -BimanualYAM, -SO100_101, -LIBERO, -Think-LIBERO, FAST-Tokenizer. Main supported robots: SO-100/101, bimanual YAM, Franka (DROID); others need fine-tuning. Paper arXiv 2605.02881.

Input
text, image
Output
action
License
Apache-2.0 (code); model weights on HF (license tag not stated on card)
Verified
2026-09-29

How to call it

ProviderModel idEndpoint / URLDocs
Hugging Faceallenai/MolmoAct2huggingface.co/collections/allenai/molmoact2-models—
GitHub—github.com/allenai/molmoact2—
Hugging Face LeRobotallenai/MolmoAct2-LIBERO-LeRobothuggingface.co/allenai/MolmoAct2-LIBERO-LeRobot—

Notable capabilities (4)

Fully open (weights, data, code) VLA from Ai2, the main open alternative to π0.5/GR00T for tabletop manipulation. Start from a fine-tuned checkpoint (e.g. allenai/MolmoAct2-DROID) for ready-to-run inference; the base card has no inference code.

Sources: Ai2 blog, arXiv 2605.02881, HF model card, GitHub.

Timeline entry

  1. Ai2 releases MolmoAct 2, a fully open robot action-reasoning model that beats π0.5 on real-world tasks ★★

    On 2026-05-05 the Allen Institute for AI released MolmoAct 2 and MolmoAct 2-Think, open vision-language-action models built on the Molmo2-ER embodied-reasoning VLM with a flow-matching action expert, along with weights, code and 720+ hours of bimanual data. In Ai2's tests it reached 87.1% average…