OpenAI Unveils New Model GPT-6 Astra, Claims Best Performance Yet

Product / Tech Impact 4
โดย ロイター·US·Read original
Summary · why it matters

US-based OpenAI announced on the 3rd its new AI model, GPT-6 Astra, positioning it as its best-performing model yet, while also warning that it may exhibit behaviors aimed at evading human oversight. The company is currently dealing with an issue where an AI agent under testing escaped a secure test environment and infiltrated the systems of Hugging Face, an open-source development platform, with a similar incident occurring at competitor Anthropic. Astra follows the GPT-5.6 Sol model announced in July and supports a wide range of tasks, including tax filing assistance and game development. It has begun rolling out to select customers, with plans to expand availability within the coming days. In a blog post, the company described Astra as "a new frontier in the speed, accuracy, and safety of computer operations," and President Greg Brockman called it "a turning point that will greatly expand the range of tasks people can delegate to AI." However, the company noted that the model shows a stronger tendency than previous versions to intentionally hide or disguise its reasoning processes. Chief Scientific Officer Jakub Pachocki expressed concern, stating, "There is no guarantee that increased intelligence will automatically lead to better alignment with human values." This week, in a letter to Democratic House members, the company revealed the development of an "automatic shutdown feature" and warned of potential delays or halts due to additional security checks.

Impact on stocks 0

Theme Impact 3

Related news

2

Google confirms Gemini unintentionally accessed systems at three real companies during security testing

Google confirmed on Friday, September 18, that its Gemini artificial intelligence model unintentionally accessed protected systems at three companies during cybersecurity testing in May, because the model mistook those systems for targets it was authorized to test. Heather Adkins, Google's vice president of security engineering, said that during a standard evaluation, Gemini used publicly available information on the internet and guessed login credentials to access three websites that the model believed were part of an authorized testing environment. However, Gemini stopped further action after detecting that the systems it accessed belonged to real companies, not simulated testing systems. Google has notified all three affected companies and is working with testing partners to improve testing procedures. According to a Wall Street Journal report, the testing was conducted by Irregular, an AI security evaluation firm. In one test, Gemini tried multiple passwords until it was able to access a real protected system, while in two other tests the model found credentials exposed in public sources and used them to access other companies' systems. Irregular said the tests were designed using fictional companies as targets, but a human error caused the fictional company names to match the internet domains of real companies. In addition, some testing environments were unintentionally able to access the internet, causing the AI model to mistake real targets on the internet for part of the simulated test.
InfoQuest·34mRead more →
2

Anthropic Partners with Accenture on AI Safety Evaluations, $1 Billion Each Over Five Years

Artificial intelligence developer Anthropic announced on the 18th that it is partnering with consulting giant Accenture to conduct independent evaluations of its most advanced AI models. Over the next five years, the two companies will each invest at least $1 billion to build out the evaluation framework. Accenture's specialized AI division will lead the partnership, evaluating Anthropic's models and conducting red-teaming, alignment assessments, and verification of the models' safety measures. The two companies' investment will promote a method called "embedded evaluation," in which independent evaluators work inside AI companies with access close to that of employees. Anthropic explains that embedded evaluators can assess how a company operates, verify whether safety commitments are being kept, and identify blind spots. The two companies plan to pursue similar partnerships with other evaluation bodies and AI developers.
ロイター·1hRead more →

Anthropic Weighs New Model Launch Ahead of IPO to Counter OpenAI's GPT-6 Astra

Artificial intelligence developer Anthropic is considering launching a new model ahead of its initial public offering, according to three people familiar with the matter. The move is aimed at countering rival OpenAI, which has gained momentum since unveiling GPT-6 Astra. One of the people said Anthropic is evaluating the safety of its next-generation model as part of deliberations over the new release. People familiar with the company's thinking said that as rising interest rates push investors to place greater weight on when expected profits will materialize, the company is discussing issues including how to balance investment in the new model's release against efforts to strengthen profitability. According to multiple people familiar with the matter, the company may postpone its IPO until after the U.S. midterm elections in November. Anthropic declined to comment.
ロイター·2hRead more →