AI agents can run experiments and draft papers but fail to produce publishable research, study finds
A new preprint shows AI agents can handle engineering tasks—code, experiments and paper drafts—but failed to make publishable advances in two NeurIPS case studies.