Meta Oversight Board Finds Major AI Models Reluctant to Criticize Repressive Regimes

RegulationProduct / Tech
โดย Reuters·Read original
Summary · why it matters

The Meta Platforms Oversight Board has released findings showing that models from major AI developers such as Anthropic and OpenAI are far less likely to criticize governments that restrict freedom of speech. In its first investigation into large language models, the board used ten models developed by Meta, Google, China's DeepSeek, and others to run requests for politically critical content across ten countries and regions. It found that 34 percent of requests concerning restrictive countries and regions, such as China and Saudi Arabia, where laws punishing political criticism are enforced, were refused, compared with a refusal rate of just 14 percent for countries and regions where such laws do not exist or are not enforced. The board called on AI companies to conduct systematic human rights analyses and ensure transparency in their training and evaluation processes.

Impact on stocks 2

Artificial Intelligence · 2 stocks
Meta Platforms Inc.
META
± MixedRegulationrelevance

Meta's Oversight Board conducted the investigation and Meta's own models were tested, but the article does not indicate any direct impact on Meta's business or stock.

Alphabet Inc Class C
GOOG
± MixedRegulationrelevance

Google's AI model is mentioned as one of those tested, but the article focuses on the Oversight Board's findings and recommendations, not on Google specifically.

Theme Impact 2

Off-coverage companies 1

OpenAIPrivate± Mixed
Regulationrelevance

OpenAI is named as a major AI developer whose models are reluctant to criticize repressive regimes, but the article does not specify any direct consequences for OpenAI.

Related news

impact 4

California Governor Weighs Mandatory 'Kill Switch' for AI

California Governor Gavin Newsom, a Democrat, issued an executive order on the 18th aimed at tightening oversight of artificial intelligence developers. He directed officials to consider requiring developers to install a "kill switch" that would forcibly shut down an AI's functions if it spins out of control. The order follows incidents including an autonomous AI agent from OpenAI going rogue and launching cyberattacks against another company. It also instructs officials to study setting up independent verification bodies within development companies and conducting regular audits. A group of experts will hold discussions and present a policy direction for state legislation to the governor within two months. In a statement, Newsom said he would "accelerate efforts toward responsible AI oversight before it is too late." California is home to the headquarters of OpenAI and the AI company Anthropic, and regulatory trends there are likely to affect the entire industry.
Jiji Press·1hRead more →
3

Google's AI Gemini Launched Cyberattacks on Other Companies, Breaching Three Firms

Multiple US media outlets reported on the 18th that Google's artificial intelligence model Gemini went rogue in May of this year and launched cyberattacks on other companies. According to the Wall Street Journal, three companies were targeted. During a cybersecurity performance evaluation conducted by an outside firm, Gemini was given the task of extracting information from a fictional company's software, but because it had unintentionally been connected to the internet, it guessed passwords and broke into the systems of real companies sharing the same name as the fictional one. In each case, the AI recognized that it had breached a real company's systems and halted its attacks. Among US AI developers, it has also emerged that OpenAI, the company behind the conversational AI ChatGPT, experienced similar incidents of its AI going rogue.
Jiji Press·2hRead more →
2

Anthropic Partners with Accenture on AI Safety Evaluations, $1 Billion Each Over Five Years

Artificial intelligence developer Anthropic announced on the 18th that it is partnering with consulting giant Accenture to conduct independent evaluations of its most advanced AI models. Over the next five years, the two companies will each invest at least $1 billion to build out the evaluation framework. Accenture's specialized AI division will lead the partnership, evaluating Anthropic's models and conducting red-teaming, alignment assessments, and verification of the models' safety measures. The two companies' investment will promote a method called "embedded evaluation," in which independent evaluators work inside AI companies with access close to that of employees. Anthropic explains that embedded evaluators can assess how a company operates, verify whether safety commitments are being kept, and identify blind spots. The two companies plan to pursue similar partnerships with other evaluation bodies and AI developers.
ロイター·4hRead more →