AI Agents
Agents Ran Hundreds of Experiments, Wrote the Paper, and Got 2 Out of 6
Princeton and the UK AI Security Institute gave an agent the central question of an unpublished NeurIPS submission, then had the paper's real authors grade the result as reviewers. Both attempts were unambiguous rejections.
3h ago
