| |
A new study from Princeton researchers found that AI agents can handle the engineering tasks needed for AI research but lack the judgment and creativity to produce original research at the level of top-tier conferences, suggesting that timelines for recursive self-improvement through AI may be overly optimistic. When tested on open-ended research questions from unpublished NeurIPS papers, AI agents like Claude Opus produced work that was rejected by the original authors, demonstrating they cannot yet match human researchers' ability to choose hypotheses, evaluate evidence, and make creative decisions essential to advancing AI.
Read Full Article →
← More Tech news