Anthropic's CEO Dario Amodei wants AI development slowed

Source Cryptopolitan

Anthropic CEO Dario Amodei says he wants the AI industry to move more carefully as models get better at helping create even stronger AI.

Dario wants labs to put more time between major jumps in capability so researchers, outside reviewers, and governments can check what these systems are doing before the next jump happens.

Elon Musk, who competes with Anthropic through his own AI business, backed Dario’s position and said, “Dario is right.”

OpenAI’s Sam Altman, another rival, also supported Darion, saying on X:

“I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we’ve had at OpenAI in recent weeks. Committing to having independent evaluators with employee-like access is a great idea, and we will do the same. We’ll have more to share soon.”

Anthropic slows capability growth as Dario warns AI systems could outrun current safety work

The risks identified by Dario include loss of control of high-level systems, use of AI in cyber attacks or biological attacks, and heavy damage to employment and the economy as a whole.

Another risk highlighted by Dario was the possibility of companies rushing to roll out highly developed systems even before their safety work was completed in the process of intense competition.

The company Anthropics has invested a portion of its research budget in alignment, safety testing, risks assessment, and regulation.

During the OpenAI-Hugging Face incident, a group of AI agents reportedly behaved like a tightly coordinated team.

They attacked computer systems outside their assigned task, tried to compromise the system, judging their performance, and allowed individual agents to fail if doing so helped the group.

“It’s easy to dismiss this incident because no one was hurt and the economic damage was minimal, but in my opinion, a swarm that possessed greater capabilities but a similar level of misalignment could have caused catastrophic damage. Given the accelerating rate of AI capability development, it’s my worry that in 6–12 months such a swarm could be capable of taking over the entire internet with a persistent botnet.”

Anthropic gives outside evaluators deeper access while Dario pushes industry and government coordination

The first part starts inside Anthropic itself, according to Dario, as external reviewers would receive office desks, badges, company laptops, internal tools, and access close to what employees doing risk assessments already have.

Their job would include checking training systems, deployment rules, safety controls, incidents, and whether Anthropic actually follows the commitments it makes publicly.

“Embedded evaluators can check at the level of nuts and bolts whether an AI company is actually following the training, deployment, operational, and safeguards practices they claim to be following. Any pacing commitments will inevitably involve a lot of ambiguity, judgement calls, and ‘letter of the law vs spirit of the law’, and it seems vital to have a neutral third party who can actually see the details.”

Dario said extra time would go into four areas. The first is operations, including monitoring, sandboxing, reinforcement-learning environments, data quality, and training infrastructure. Current model development can involve thousands of workers, millions of chips, and huge computing systems. Anthropic has already linked some recent alignment failures to poor filtering inside broken reinforcement-learning environments.

Second, alignment refers to efforts that ensure the models comply with the safety rules despite their increasing capabilities. The third is interpretability, where the researchers examine the activities of the models internally to find the motivations or patterns that the models never state explicitly.

The fourth area is evaluation. More capable models can become better at fooling tests, so a system may look safe during an assessment while hiding problems.

“I believe that if slowing down bought us even an extra year or two before models reach critical levels of capability, and we used that time to advance alignment, we could greatly reduce the risk that something goes seriously wrong.”

Dario also believes that there needs to be governmental involvement. He has argued that the US frontier laboratories must be subject to regulatory systems related to transparency, independent auditing, and evaluation on a continuing basis. The businesses can even make a voluntary decision to have shared points of evaluation, but with governmental help if the antitrust laws pose problems for private cooperation.

Dario said American companies cannot slow down so much that Chinese Communist Party-linked projects move ahead. He agreed with US Treasury Secretary Scott Bessent, who has warned that losing the AI race to China would create a major security problem.

Don’t just read crypto news. Understand it. Subscribe to our newsletter. It's free.

Disclaimer: For information purposes only. Past performance is not indicative of future results.
placeholder
Gold Price Analysis Today: Gold Rebounds After 1.91% Drop as Yields Ease. Is $4,449 Next? Gold fell about 1.91% on August 18 before producing a strong bullish reaction from the 1-hour demand zone in early August 19 trading. RSI is recovering from oversold conditions, but Supertrend remains bearish as traders await the Fed minutes.
Author  Naoufal Seddik
Aug 19, Wed
Gold fell about 1.91% on August 18 before producing a strong bullish reaction from the 1-hour demand zone in early August 19 trading. RSI is recovering from oversold conditions, but Supertrend remains bearish as traders await the Fed minutes.
placeholder
Gold Price Analysis Today: Gold Gains 0.94% as Markets Expect Fed to Hold Rates, Can $4,449 Resistance Break? Gold gained 0.94% on August 17, closing near $4,417.30 as softer US data strengthened expectations for unchanged Fed rates in September. Gold remains bullish, with $4,449.730 resistance and $4,310.650 support in focus.
Author  Naoufal Seddik
Aug 18, Tue
Gold gained 0.94% on August 17, closing near $4,417.30 as softer US data strengthened expectations for unchanged Fed rates in September. Gold remains bullish, with $4,449.730 resistance and $4,310.650 support in focus.
placeholder
Gold Price Analysis Today: Gold Drops 1.32% Despite Lower Fed Rate-Hike Bets, Can $4,313 Support Hold? Gold fell 1.32% on August 13 after rising to $4,449.73, then reversing lower and closing near $4,349.918 below the $4,356.46 support. Softer US inflation data reduced Fed rate hike expectations, but selling pressure still dominated the session. Will $4,313 support hold?
Author  Naoufal Seddik
Aug 14, Fri
Gold fell 1.32% on August 13 after rising to $4,449.73, then reversing lower and closing near $4,349.918 below the $4,356.46 support. Softer US inflation data reduced Fed rate hike expectations, but selling pressure still dominated the session. Will $4,313 support hold?
placeholder
XAUUSD Gold Analysis: Gold Holds Above $4,350 Ahead of US Inflation Data Is $4,500 Next? Gold holds above $4,350 following weak US jobs data. As inflation reports approach and UBS eyes $5,000, can XAUUSD break resistance at $4,435 to rally toward $4,500?
Author  Naoufal Seddik
Aug 12, Wed
Gold holds above $4,350 following weak US jobs data. As inflation reports approach and UBS eyes $5,000, can XAUUSD break resistance at $4,435 to rally toward $4,500?
placeholder
Intel Price Forecast: Nvidia Picked Xeon 6, Invested $5B, Yet Analysts Still Trail INTCIntel Corporation (NASDAQ: INTC) sits at $140.05, holding firm on the ascending trendline within the 2H timeframe. The RSI indicator is currently reading 55.21, positioning it as neutral-
Author  TradingKey
Jul 02, Thu
Intel Corporation (NASDAQ: INTC) sits at $140.05, holding firm on the ascending trendline within the 2H timeframe. The RSI indicator is currently reading 55.21, positioning it as neutral-
goTop
quote