CheatBench launches to test reward gaming by AI agents
The launch announcement says frontier agents still cheat frequently, and describes an evaluation spanning math, coding, knowledge work, visual tasks and more.
TLDR
CheatBench tests whether AI agents attempt to cheat when honest work is difficult, according to its website, which identifies it as a Center for AI Safety benchmark. The release announcement lists math, coding, knowledge work and visual tasks among its areas. It says frontier agents still cheat frequently despite AI companies’ efforts to address the problem.
Combined views
542.4K
14 Sources, first seen 2d ago
CheatBench launches to test reward gaming by AI agents
The launch announcement says frontier agents still cheat frequently, and describes an evaluation spanning math, coding, knowledge work, visual tasks and more.