π₯ Explore this insightful post from Hacker News π
π **Category**:
π **What Youβll Learn**:
Anthropic, a company founded by OpenAI exiles worried about the dangers of AI, is loosening its core safety principle in response to competition.
Instead of self-imposed guardrails constraining its development of AI models, Anthropic is adopting a nonbinding safety framework that it says can and will change.
In a blog post Tuesday outlining its new policy, Anthropic said shortcomings in its two-year-old Responsible Scaling Policy could hinder its ability to compete in a rapidly growing AI market.
The announcement is surprising, because Anthropic has described itself as the AI company with a βsoul.β It also comes the same week that Anthropic is fighting a significant battle with the Pentagon over AI red lines.
The policy change is separate and unrelated to Anthropicβs discussions with the Pentagon, according to a source familiar with the matter. Defense Secretary Pete Hegseth gave Anthropic CEO Dario Amodei an ultimatum on Tuesday to roll back the companyβs AI safeguards or risk losing a $200 million Pentagon contract. The Pentagon threatened to put Anthropic on what is effectively a government blacklist.
But the company said in its blog post that its previous safety policy was designed to build industry consensus around mitigating AI risks β guardrails that the industry blew through. Anthropic also noted its safety policy was out of step with Washingtonβs current anti-regulatory political climate.
Anthropicβs previous policy stipulated that it should pause training more powerful models if their capabilities outstripped the companyβs ability to control them and ensure their safety β a measure thatβs been removed in the new policy. Anthropic argued that responsible AI developers pausing growth while less careful actors plowed ahead could βresult in a world that is less safe.β
As part of the new policy, Anthropic said it will separate its own safety plans from its recommendations for the AI industry.
Anthropic wrote that it had hoped its original safety principles βwould encourage other AI companies to introduce similar policies. This is the idea of a βrace to the topβ (the converse of a βrace to the bottomβ), in which different industry players are incentivized to improve, rather than weaken, their modelsβ safeguards and their overall safety posture.β
The company now suggests that hasnβt played out.
In a statement to CNN, an Anthropic spokesperson described the updated policy as βthe strongest to date on the level of public accountability and transparency.β
βWeβve gone a significant step further from our prior policies by committing to publicly publish detailed reports at regular intervals on our plans to strengthen our risk mitigations, as well as the threat models and capabilities of all our models,β the statement said. βFrom the beginning, weβve said the pace of AI and uncertainties in the field would require us to rapidly iterate and improve the policy.β
Anthropicβs new safety policy includes a βFrontier Safety Roadmapβ that outlines the companyβs self-imposed guidelines and safeguards. But the company acknowledged the new framework is more flexible than its past policy.
βRather than being hard commitments, these are public goals that we will openly grade our progress towards,β the company said in its blog post.
The change comes a day after Defense Secretary Pete Hegseth gave Anthropic CEO Dario Amodei a Friday deadline to roll back the companyβs AI safeguards, or risk losing a $200 million Pentagon contract and being put on what is effectively a government blacklist.
Anthropic has concerns over two issues that it isnβt willing to drop, according to a source familiar with the companyβs meeting with Hegseth: AI-controlled weapons and mass domestic surveillance of American citizens. Anthropic believes AI is not reliable enough to operate weapons, and there are no laws or regulations yet that cover how AI could be used in mass surveillance, a source said.
AI researchers applauded Anthropicβs stance on social media on Tuesday and expressed concerns about the idea of AI being used for government surveillance.
The company has long positioned itself as the AI business that prioritizes safety. Anthropic has published research showing how its own AI models could be capable of blackmail under certain conditions. The company recently donated $20 million to Public First Action, a political group pushing for AI safeguards and education.
But the company has faced increasing pressure and competition from both the government and its rivals. Hegseth, for example, plans to invoke the Defense Production Act on Anthropic and designate the company a supply chain risk if it does not comply with the Pentagonβs demands, CNN reported on Tuesday. OpenAI and Anthropic have also been locked in a race to launch new enterprise AI tools in a bid to win the workplace.
Jared Kaplan, Anthropicβs chief science officer, suggested in an interview with Time that the change was made in the name of safety more than increased competition.
βWe felt that it wouldnβt actually help anyone for us to stop training AI models,β Kaplan told the magazine. βWe didnβt really feel, with the rapid advance of AI, that it made sense for us to make unilateral commitments β¦ if competitors are blazing ahead.β
CNNβs Hadas Gold contributed to this story.
This story has been updated with additional information.
π¬ **Whatβs your take?**
Share your thoughts in the comments below!
#οΈβ£ **#Anthropic #ditches #core #safety #promise #middle #red #line #fight #Pentagon**
π **Posted on**: 1772113632
π **Want more?** Click here for more info! π
