Skip to content
Türkçe

Research · Models · Anthropic

In C5R’s AI-run lab, the best model completed only 45% of the tasks

Published: 2 sourcesTürkçe

On September 24 San Francisco-based C5R unveiled Facility-0, a lab built in twelve weeks that combines biology, chemistry and materials science, along with SciUniverse, a benchmark that measures how well models can carry out physical experiments. Frontier models scored between 9.4% and 45.3%, with Anthropic’s Claude Fable 5.1 on top at 45.3%. The failures were recurring: trying to pipette frozen samples, reusing pipette tips across DNA-containing wells, vortexing open plates and misreading noise in spectroscopy data. The takeaway: models know the science but lack an intuition for the physical reality of a lab.

Sources

  1. C5R, post on X (@c5rcorp), (x.com)
  2. C5R, “Can frontier models carry out scientific work?” (c5r.net)

About this story

This story was posted on Instagram by @jarrus.tech on Sept. 29, 2026.

Spotted an error in this story? [email protected] · Instagram

This story in Turkish: C5R’ın yapay zekanın yönettiği laboratuvarında en iyi model görevlerin yalnızca %45’ini başardı

On the same topic