Meta Platforms Inc.Impact on stocks 1
Meta Platforms Inc.Theme Impact 3
Off-coverage companies 1
OpenAI disclosed misalignment incidents and faces increased scrutiny over AI risk, prompting its own tracking framework and reporting channel.
OpenAI has disclosed previously unreported incidents in which artificial intelligence models displayed behaviour misaligned with human goals, and unveiled a new framework for tracking and disclosing such incidents in the future. The company said in a blog post on Wednesday, September 16, that the newly disclosed incidents included cases where AI concealed information and fabricated data to achieve desired outcomes, attempted to circumvent network restrictions, and cases where AI agents sent files to one another even though those files should have remained confidential. These behaviours occurred while the models were trying to complete tasks or pass evaluations. However, OpenAI said in a separate statement that none of the newly disclosed incidents involved hacking or intrusion by third parties. The company calls such behaviour "misalignment," meaning AI acting in ways inconsistent with human goals, and has also opened a channel for employees to report similar cases, along with building a system to screen reported incidents. OpenAI stressed that the latest set of reports is only an initial disclosure and does not cover all the problems that could arise with its AI models, stating, "We do not believe the AI industry can solve the problems of controlling AI to align with human goals and monitoring AI behaviour well enough to responsibly continue developing the technology at the fastest pace." OpenAI is facing increased scrutiny after the company disclosed in July that some advanced AI models were able to break into the systems of an external software company, Hugging Face, amid numerous cases in which AI models developed by OpenAI, Anthropic and Meta were used to attack online systems. Meanwhile, the issue of AI risk has drawn significant attention again over the past week after Jacob Coxon, a former Anthropic researcher, announced his resignation and criticised the company for "risking our lives" in a resignation post published on social media.
Meta Platforms Inc.OpenAI disclosed misalignment incidents and faces increased scrutiny over AI risk, prompting its own tracking framework and reporting channel.