Technology·

Current AI Agents Fall Short of Autonomous Scientific Research, Studies Show

Recent evaluations by Epoch AI and Anthropic reveal that advanced AI models struggle with genuine creative thinking and critical self-assessment during scientific experiments. Scoring low compared to human benchmarks, these models frequently overstate their findings and lack the autonomy required for independent research, highlighting a critical gap in current artificial intelligence capabilities.

Source: The Decoder