Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. OpenAI releases LifeSciBench, 750 expert-written…

OpenAI releases LifeSciBench, 750 expert-written life-science research tasks; GPT-Rosalind leads with a 36% pass rate

★★benchmarkOpenAIconfidence: medium

On June 17, 2026 OpenAI introduced LifeSciBench, a benchmark of 750 expert-authored tasks that span seven research workflows and seven biological domains, graded with 19,020 rubric criteria. Its life-sciences model GPT-Rosalind scored best but passed only 36.1% of tasks, and 22.8% of tasks were passed by no model, so the benchmark is far from saturated.

Key facts

What happened

OpenAI released LifeSciBench, with a preprint, as a benchmark of realistic life-science research work rather than exam questions. It was published two months after GPT-Rosalind, which led the results.

Why it matters

It is a large, expert-graded measure of AI for biology research. Its main finding is that models still struggle to read real scientific data files, the skill research agents need most. As OpenAI built the benchmark and its own model leads it, independent replication matters.

Changelog

  • 2026-10-04: created (resolves leads.md line)

Related events

  1. OpenAI launches GPT-Rosalind, a trusted-access reasoning model for life-sciences research ★★★
  2. BixBench3 tests AI agents on whole computational-biology studies from raw data; best score 0.48 (GPT-5.6 Sol) ★★

Sources (4)

id: 2026-06-17-openai-lifescibench · updated 2026-10-04 · open in the interactive timeline