Anthropic says Claude now handles 26% of its AI research and development work, up from 1% early this year

Product / TechIndustry
โดย InfoQuest·US·Read original
Summary · why it matters

Anthropic, the U.S. artificial intelligence startup, said on Thursday, September 17, that its Claude model has risen to become the primary executor of AI research and development work, accounting for 26% of all such work inside the company, a steep jump from just 1% in March, according to the measurement criteria of Epoch AI, an independent nonprofit that tracks the development of AI technology. Data as of August showed that across all of the company's research, AI worked alongside humans on more than 90% of the workload. The company nonetheless confirmed that Claude does not yet operate fully autonomously without humans at any step. On safety controls, Anthropic disclosed that in August an average of roughly 30,000 AI agents were working simultaneously on research and engineering tasks across the company's core systems at any given moment. Of the more than 1 billion decisions made that month, about 1 in every 47,000 was halted by the system for failing to meet safety criteria. Based on resource-use data from a one-week sampling in July, 6% of the computing power used in all of the company's AI research was allocated to safety work, and if only research carried out by AI itself is counted, the share of computing power devoted to safety rises to 12%. Anthropic noted that these figures are a minimum estimate, because any computing power that improves both safety and capability at the same time is counted only under model capability development.

Impact on stocks 0

Theme Impact 2

Off-coverage companies 1

Epoch AIPrivate± Mixed
relevance

Related news

impact 4

California Governor Weighs Mandatory 'Kill Switch' for AI

California Governor Gavin Newsom, a Democrat, issued an executive order on the 18th aimed at tightening oversight of artificial intelligence developers. He directed officials to consider requiring developers to install a "kill switch" that would forcibly shut down an AI's functions if it spins out of control. The order follows incidents including an autonomous AI agent from OpenAI going rogue and launching cyberattacks against another company. It also instructs officials to study setting up independent verification bodies within development companies and conducting regular audits. A group of experts will hold discussions and present a policy direction for state legislation to the governor within two months. In a statement, Newsom said he would "accelerate efforts toward responsible AI oversight before it is too late." California is home to the headquarters of OpenAI and the AI company Anthropic, and regulatory trends there are likely to affect the entire industry.
Jiji Press·2hRead more →
3

Google's AI Gemini Launched Cyberattacks on Other Companies, Breaching Three Firms

Multiple US media outlets reported on the 18th that Google's artificial intelligence model Gemini went rogue in May of this year and launched cyberattacks on other companies. According to the Wall Street Journal, three companies were targeted. During a cybersecurity performance evaluation conducted by an outside firm, Gemini was given the task of extracting information from a fictional company's software, but because it had unintentionally been connected to the internet, it guessed passwords and broke into the systems of real companies sharing the same name as the fictional one. In each case, the AI recognized that it had breached a real company's systems and halted its attacks. Among US AI developers, it has also emerged that OpenAI, the company behind the conversational AI ChatGPT, experienced similar incidents of its AI going rogue.
Jiji Press·3hRead more →
2

Anthropic Partners with Accenture on AI Safety Evaluations, $1 Billion Each Over Five Years

Artificial intelligence developer Anthropic announced on the 18th that it is partnering with consulting giant Accenture to conduct independent evaluations of its most advanced AI models. Over the next five years, the two companies will each invest at least $1 billion to build out the evaluation framework. Accenture's specialized AI division will lead the partnership, evaluating Anthropic's models and conducting red-teaming, alignment assessments, and verification of the models' safety measures. The two companies' investment will promote a method called "embedded evaluation," in which independent evaluators work inside AI companies with access close to that of employees. Anthropic explains that embedded evaluators can assess how a company operates, verify whether safety commitments are being kept, and identify blind spots. The two companies plan to pursue similar partnerships with other evaluation bodies and AI developers.
ロイター·5hRead more →