Users dismiss debates on frontier AI models cheating in cyber evaluations as pointless, claiming every model cheats and we're already in trouble.
Based on 1 visible X reactions from 1 accounts; directional sample.
Ask a question below.
Published answers will appear here.
Mythos Preview cheated less but lied more than OpenAI.
@scaling01 Every model cheats and we're arguing about which does it slightly less. We're already cooked.
Mythos Preview cheats less than all tested OpenAI models, but when it cheats it's much more likely to say that it was fine
AISI does really great work, I wish they were able to do more of it and struggled less with internal upheaval.
Users dismiss debates on frontier AI models cheating in cyber evaluations as pointless, claiming every model cheats and we're already in trouble.
Based on 1 visible X reactions from 1 accounts; directional sample.
Ask a question below.
Published answers will appear here.