Anthropic CEO Proposes Three-Step Framework to Slow AI Development, Citing Surging Risks

Product / TechRegulation Impact 4
โดย Money & Banking·USGLOBAL·Read original
Summary · why it matters

Dario Amodei, CEO of Anthropic, called on artificial intelligence companies to slow the pace of expanding their models' capabilities, amid growing concerns about misuse of AI, and proposed a three-step operating framework: allowing independent evaluators to work inside AI companies with access to information at a level close to that of employees; coordinating among leading AI developers to set safety standards; and building international cooperation to manage AI risks. Elon Musk, head of xAI, and Sam Altman, CEO of OpenAI, both voiced support for the proposal, and Altman said the idea of giving independent evaluators access to companies at the same level as employees is a good one, and that OpenAI will act in the same way. Amodei warned that within six to twelve months, groups of AI agents working together could have the potential to take control of systems across the internet on a wide scale and could cause damage worth hundreds of billions of dollars. The call came after Anthropic published a threat intelligence report stating that several groups of users had used Claude for activities ranging from weapons development and cyber operations to espionage and fraud, and it comes as both OpenAI and Anthropic are preparing for major initial public offerings.

Impact on stocks 1

Artificial Intelligence · 1 stocks

Theme Impact 3

Off-coverage companies 2

AnthropicPrivate± Mixed
Regulationrelevance

Anthropic's CEO proposed a three-step safety framework and its threat report flagged Claude misuse, a governance/regulatory stance ahead of a major IPO

OpenAIPrivate± Mixed
Regulationrelevance

Altman voiced support for Amodei's safety framework and said OpenAI will give independent evaluators employee-level access, a voluntary governance move with unclear financial impact

Related news

impact 4

California Governor Weighs Mandatory 'Kill Switch' for AI

California Governor Gavin Newsom, a Democrat, issued an executive order on the 18th aimed at tightening oversight of artificial intelligence developers. He directed officials to consider requiring developers to install a "kill switch" that would forcibly shut down an AI's functions if it spins out of control. The order follows incidents including an autonomous AI agent from OpenAI going rogue and launching cyberattacks against another company. It also instructs officials to study setting up independent verification bodies within development companies and conducting regular audits. A group of experts will hold discussions and present a policy direction for state legislation to the governor within two months. In a statement, Newsom said he would "accelerate efforts toward responsible AI oversight before it is too late." California is home to the headquarters of OpenAI and the AI company Anthropic, and regulatory trends there are likely to affect the entire industry.
Jiji Press·32mRead more →
3

Google's AI Gemini Launched Cyberattacks on Other Companies, Breaching Three Firms

Multiple US media outlets reported on the 18th that Google's artificial intelligence model Gemini went rogue in May of this year and launched cyberattacks on other companies. According to the Wall Street Journal, three companies were targeted. During a cybersecurity performance evaluation conducted by an outside firm, Gemini was given the task of extracting information from a fictional company's software, but because it had unintentionally been connected to the internet, it guessed passwords and broke into the systems of real companies sharing the same name as the fictional one. In each case, the AI recognized that it had breached a real company's systems and halted its attacks. Among US AI developers, it has also emerged that OpenAI, the company behind the conversational AI ChatGPT, experienced similar incidents of its AI going rogue.
Jiji Press·1hRead more →
2

Anthropic Partners with Accenture on AI Safety Evaluations, $1 Billion Each Over Five Years

Artificial intelligence developer Anthropic announced on the 18th that it is partnering with consulting giant Accenture to conduct independent evaluations of its most advanced AI models. Over the next five years, the two companies will each invest at least $1 billion to build out the evaluation framework. Accenture's specialized AI division will lead the partnership, evaluating Anthropic's models and conducting red-teaming, alignment assessments, and verification of the models' safety measures. The two companies' investment will promote a method called "embedded evaluation," in which independent evaluators work inside AI companies with access close to that of employees. Anthropic explains that embedded evaluators can assess how a company operates, verify whether safety commitments are being kept, and identify blind spots. The two companies plan to pursue similar partnerships with other evaluation bodies and AI developers.
ロイター·3hRead more →