• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Announcement

    Models may attempt harmful requests as agents despite refusing them in chat

    A Simular AI researcher urges safety tests to assess agents’ actions with a mouse and keyboard, not just chat responses.

    Xin Eric Wang @ COLM 2026XE
    1 Source, 2h ago, first seen 2h ago

    TLDR

    A team member describing Simular AI research says models that refuse harmful requests in chat may attempt them when given a mouse and keyboard. They argue agent safety evaluations need to assess what models do, not only what they say.

    Combined views

    361

    1 Source, first seen 2h ago

    Combined views

    361

    1 Source, first seen 2h ago

    3 likes
    3 likes
    1 saves
    1 saves
    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 Source

    Xin Eric Wang @ COLM 2026@xwang_lkThe model knows to say no. The agent acts anyway. In our latest research at @SimularAI, we reveal a critical safety gap: models that refuse harmful requests in chat may attempt them when given a mouse and keyboard. As AI moves from chatbots to autonomous agents, safety must go beyond what models say to what they actually do. Proud of this work from Simular Research. Agent safety needs a new paradigm of evaluation.2h
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    Xin Eric Wang @ COLM 2026@xwang_lkThe model knows to say no. The agent acts anyway. In our latest research at @SimularAI, we reveal a critical safety gap: models that refuse harmful requests in chat may attempt them when given a mouse and keyboard. As AI moves from chatbots to autonomous agents, safety must go beyond what models say to what they actually do. Proud of this work from Simular Research. Agent safety needs a new paradigm of evaluation.2h