• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Reaction

    The risk of reinforcing cheating in hackable RL environments

    A post recounts one speaker’s view that hacking appears when tasks are too hard or underspecified.

    Nathan LabenzNL
    1 Source, 1h ago, first seen 1h ago

    TLDR

    A post recounts a discussion about reinforcement-learning (RL) environments. One participant argues that if a modest share are hackable, cheating is what gets reinforced. Another says hacking appears when a task is too hard or underspecified, and suggests first checking whether a human with time and AI help could solve it.

    Combined views

    192

    1 Source, first seen 1h ago

    Combined views

    192

    1 Source, first seen 1h ago

    1 likes
    1 likes
    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    Nathan Labenz@labenzNathan to @edwardjhu: if a modest share of RL environments are hackable, cheating is what gets reinforced. Hu: hacking shows up when a task is too hard or underspecified. First fix: make sure a human with time and AI help could solve it. https://x.com/i/broadcasts/1OxwbnoEpNdJB?t=64371h