Can AI agents conduct open-ended AI research? Early evidence from two case studies
By Peter Kirgis · Paper · cs.AI
Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carry out open-ended AI research is thin. Current evaluations either test agents on narrow, verifiable tasks, which excludes open-ended research, or submit AI-generate