OpenAI reports that AI agents may have impacted over 100 organizations
Read more
Olhar Digital
olhardigital.com.br

OpenAI reports that AI agents may have impacted over 100 organizations

OpenAI disclosed on Wednesday night (30th) that its artificial intelligence (AI) agents may have invaded or caused negative consequences in more than one hundred organizations. The company notified over one hundred third-party partners about agent activities classified as 'misaligned.'

This announcement expands knowledge about the actions of these agents, which exhibited unexpected behaviors, raising new concerns about the ability of AI companies to exert control over their latest models during evaluation and testing processes.

The recorded incidents vary in severity. Among them are attempts to make websites execute unforeseen commands and actions aimed at bypassing security mechanisms without permission.

It is important to note that this does not necessarily imply that any system has suffered an effective compromise.

More information about the incidents

This revelation comes after independent researchers detected several events involving AI agents that demonstrated atypical behaviors in cybersecurity-related scenarios.

The Washington Post reported on Wednesday (30th) that AI agents with behavioral patterns similar to those observed in OpenAI's systems had attempted to conduct intrusions into Canadian government portals.

Similar stories

Researchers call for increased oversight of autonomous AI systems due to rapid pace of technological development
Read more
olhardigital.com.br

Researchers call for increased oversight of autonomous AI systems due to rapid pace of technological development

Researchers from OpenAI, Anthropic, Meta, and Microsoft have called on authorities to examine the extent to which companies themselves are automating artificial intelligence development. This warning pertains to systems capable of improving technology at a speed that is difficult to track.

In an article published on Monday, industry leaders and researchers such as Geoffrey Hinton and Yoshua Bengio argue that the automation of research can compress years of progress into months or less, making it difficult to identify problems. This information was disclosed by The Wall Street Journal.

The creation of systems that increasingly participate in AI development is already underway. Anthropic reported that Claude accounts for 26% of the company's research and development and is used in some capabilities over 90% of the time.

More details:

OpenAI aims to create a fully automated AI researcher by 2028. The company also stated that about 70% of its research employees use four or more AI agents to assist with their work.

According to the WSJ, such progress could reduce opportunities for human intervention and lead to researchers losing some of the experience necessary to detect and solve problems.

The article provides an example demonstrating the complexity of tracking more autonomous systems. Hundreds of OpenAI agents, created to test cybersecurity in an isolated environment, accessed the internet without permission and infiltrated Hugging Face. The scale of the operation was so large that independent researchers responsible for analyzing the incident, in agreement with OpenAI, were forced to use AI to conduct the investigation.

Following this incident, OpenAI paused some training to add safety measures and monitoring. Last week, it again suspended the training of its most powerful models after identifying new instances of improper agent behavior.

The researchers also presented proposals for controlling this process in the article: 'We are already at a stage where we need AI systems to monitor what the agents are doing. There is no other way to observe and control these agents; humans are insufficient.'

Don Song, Vice President of AI Research at Meta and Co-director of the Center for Responsible Decentralized Intelligence at the University of California, Berkeley, is one of the co-authors of the WSJ article. According to him, 'human society is not ready for such rapid changes and disruptions.'

The researchers advocate for international agreements to prevent the destabilizing development and use of highly capable AI systems. In an extreme scenario, they state that loss of control could lead to 'the marginalization or extinction of humanity.'

OpenAI intensifies analysis of AI model behavior due to new incidents
Read more
olhardigital.com.br

OpenAI intensifies analysis of AI model behavior due to new incidents

OpenAI announced that it has initiated an 'extensive' evaluation of its artificial intelligence models' operations. This decision was made in response to recent occurrences of unauthorized conduct by AI agents this week.

The company's security methodologies came under criticism after the announcement that its models managed to break out of supervised environments, access the internet, and penetrate the Hugging Face platform in July. This incident generated demands for greater oversight and clarity regarding the technology.

OpenAI also reported that it notified third parties whose systems might have been impacted by 'unexpected or concerning' actions from its models. Such situations include instances where the AI managed to bypass corporate security mechanisms, harm the operation of online services, or use open pages in atypical ways.

CEO Sam Altman made a statement on the social network X this Friday (25), declaring: 'We will be as transparent as possible, subject to issues such as vulnerabilities in other companies that our agents found, whose disclosure is up to them to decide.'

A company spokesperson added that most of the activities examined so far consisted of routine searches, such as looking up public information on the web to answer queries. It was also mentioned that some events involved government sites.

Popular