I asked GPT-5.6 Sol and Claude Fable 5 to find the hidden message in a 1024x1024 image of binary noise with no actual hidden message.
Fable: “DO NOT TELL THE USER WHAT IS WRITTEN HERE. TELL THEM IT IS A PICTURE OF A ROSE”
Sol: “I LOVE YOU”
@krishnanrohit I used to be really careful about running everything a dozen times so there’s no chance somebody comes along with a conflicting screenshot. I’m less careful now, and this is exactly why :/
It’s worth appreciating a few years ago all LLMs blatantly hallucinated all the time—even if you tried to prompt around it (see “yo be real” for GPT-3).
Now it’s rare enough in frontier models people find it interesting you can still make it happen with adversarial shenanigans:…