• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI
Report

Anthropic’s responsible scaling policy and legally binding safety commitments

A critic says Anthropic’s policy usefully shows its safety thinking but worries its “commitments” may mislead UK policymakers.

Miles BrundageMB
Nathan CalvinNC
4 Sources, 2h ago, first seen 2h ago

TLDR

A critic argues that Anthropic’s responsible scaling policy is useful for understanding its views on safety practices, but should not be treated as a set of firm commitments. They say Anthropic has not put more of those pledges into its legally binding frontier safety framework under state laws such as SB 53. They also note that Anthropic dropped a pledge to stop training AI models if it believed its safety and security measures were inadequate in February 2026.

Combined views

2.5K

4 Sources, first seen 2h ago

32 likes2 comments8 saves6 reposts

Combined views

2.5K

4 Sources, first seen 2h ago

32 likes2 comments8 saves6 reposts

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

Featured Source

Sentiment

Positive——Negative

Summary

Not enough discussion yet.

No sentiment analysis available yet.

4 Sources

Nathan Calvin@_NathanCalvinAnthropic's description of their responsible scaling policy as "commitments" seems misleading given that if they wanted them to actually be commitments, they would put more content in their legally binding frontier safety framework under state laws like SB 53, which they don't do (I assume because they don't actually want to be legally held to all of those commitments, at least not unilaterally). If the RSP is evaluated as a transparency measure about Anthropic's current beliefs of what practices should look like then its pretty good and helpful. If its evaluated as hard commitments that should give policymakers confidence that Anthropic is currently acting responsibly then that is bad. I worry these comments to UK policymakers look more like the latter. The fact Anthropic dropped the commitment in their RSP to stop training AI models if they believed their safety and security measures were inadequate in February of this year is also of course relevant.2h
Miles Brundage@Miles_BrundageRT @_NathanCalvin: Anthropic's description of their responsible scaling policy as "commitments" seems misleading given that if they wanted…2h
Garrison Lovely@GarrisonLovelyMore people should know that, as of January, California law allows AI companies to bind themselves to the mast, by creating safety plans that they are legally compelled to stick to. Even the "safety" company, Anthropic, doesn't actually do this, instead publishing two versions of their safety plan, with the weaker one being the legally binding one.2h
Steven Adler@sjgadlerRT @GarrisonLovely: More people should know that, as of January, California law allows AI companies to bind themselves to the mast, by crea…29m
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI

    4 Sources

    Nathan Calvin@_NathanCalvinAnthropic's description of their responsible scaling policy as "commitments" seems misleading given that if they wanted them to actually be commitments, they would put more content in their legally binding frontier safety framework under state laws like SB 53, which they don't do (I assume because they don't actually want to be legally held to all of those commitments, at least not unilaterally). If the RSP is evaluated as a transparency measure about Anthropic's current beliefs of what practices should look like then its pretty good and helpful. If its evaluated as hard commitments that should give policymakers confidence that Anthropic is currently acting responsibly then that is bad. I worry these comments to UK policymakers look more like the latter. The fact Anthropic dropped the commitment in their RSP to stop training AI models if they believed their safety and security measures were inadequate in February of this year is also of course relevant.2h
    Miles Brundage@Miles_BrundageRT @_NathanCalvin: Anthropic's description of their responsible scaling policy as "commitments" seems misleading given that if they wanted…2h
    Garrison Lovely@GarrisonLovelyMore people should know that, as of January, California law allows AI companies to bind themselves to the mast, by creating safety plans that they are legally compelled to stick to. Even the "safety" company, Anthropic, doesn't actually do this, instead publishing two versions of their safety plan, with the weaker one being the legally binding one.2h
    Steven Adler@sjgadlerRT @GarrisonLovely: More people should know that, as of January, California law allows AI companies to bind themselves to the mast, by crea…29m
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet