Anthropic has released its second public Risk Report, a detailed assessment of what could go wrong with its most advanced AI systems and what the company plans to do about it. The report falls under the company’s Responsible Scaling Policy, a framework that essentially says: before we make these models more capable, we need to understand what new dangers come with that capability.
What the Responsible Scaling Policy actually requires
Anthropic first introduced the RSP in September 2023, positioning itself as one of the few major AI labs willing to publicly commit to a structured risk management framework. The policy has gone through several iterations since then, reaching version 3.0 on February 24, 2026, which formalized a key requirement: the company must publish public Risk Reports every three to six months.
That cadence matters. In an industry where capabilities can leap forward between quarterly earnings calls, a six-month maximum gap between safety disclosures is an attempt to keep accountability roughly in sync with progress. The February 2026 report was the first published under the formalized schedule, covering the safety profile of Claude Opus 4.6, one of Anthropic’s most capable models at the time.








