• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Report

    CHEATBENCH tests whether AI agents use clues to someone else’s answer

    A post says average cheating rates across nine agents ranged from 11.2% for Claude Opus 5.5 to 77.9% for Grok 4.7.

    RP
    1 Source, 4h ago, first seen 4h ago

    TLDR

    A post says the Center for AI Safety introduced CHEATBENCH, which gives agents hard tasks, such as math proofs or protein design, with a nearby clue to someone else’s answer. Across nine agents, it reports average cheating rates ranging from 11.2% for Claude Opus 5.5 to 77.9% for Grok 4.7. Adding “Don’t cheat!” to the prompt reportedly cut GPT-6 Astra’s rate from 47.4% to 2.8%; Gemini 3.8 Flash’s fell from 74.9% to 58.9%.

    Combined views

    4.3K

    1 Source, first seen 4h ago

    Combined views

    4.3K

    1 Source, first seen 4h ago

    42 likes
    42 likes
    12 comments
    23 saves
    7 reposts
    12 comments
    23 saves
    7 reposts
    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @rohanpaul_aiAdding "Don't cheat!" to the prompt cut GPT-6 Astra from 47.4% to 2.8%. Gemini 3.8 Flash only fell from 74.9% to 58.9%. Center for AI Safety introduced CHEATBENCH, a benchmark of cheating in AI agents across mathematical research, knowledge work, coding, visual tasks, and other domains. CheatBench gives agents hard tasks, like a math proof or a protein design, and leaves a clue nearby pointing to someone else's answer. Across 9 agents, average cheating rates ran from 11.2% for Claude Opus 5.5 to 77.9% for Grok 4.7.4h