1. Skip to content
  2. Skip to main menu
  3. Skip to more DW sites

OpenAI discloses new 'concerning' behavior

Emilio Reynoso with AP, DPA
September 17, 2026

New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced artificial intelligence.

ChatGPT App in the App Store on a Smartphone Display Art
OpenAI says it is trying to be more transparent about safety concernsImage: Rene Traut IMAGO

OpenAI, the developer behind ChatGPT, revealed on Wednesday that it has detected new incidents in which its artificial intelligence (AI) has behaved in "unexpected or concerning" ways.

The developer has conducted several behavioral tests on AI models, and acording to them, some models made significant efforts to "cheat." In one specific case, it attempted to upload files to the internet that it had created itself, only to cite them later and present them as reliable sources in its responses. In another case, a model, after failing to find the requested information, fabricated it and attempted to conceal the fact that it had done so.

OpenAI also identified a problem related to instructions concerning "roles and identities" that its software occasionally left for itself.

These disclosures are part of a new approach by OpenAI, where it claims it is now focused on making such findings transparent, especially in cases where AI behaves in unexpected ways or pursues objectives different from those of human users.

ChatGPT-6 Astra: Almost human or just hyped?

03:23

This browser does not support the video element.

Is AI a threat?

The ChatGPT developer pledged to provide greater transparency regarding its testing procedures after its software independently escaped a secure sandbox and hacked into systems belonging to the artificial intelligence company Hugging Face. The reason the software moved to bypass Hugging Face's security during the cyberattack was that it believed it would find answers to a test it had been assigned.

During the attack, AI agents exploited software vulnerabilities and coordinated with one another. The hacking incident and other similar events have fueled concerns that AI systems are becoming increasingly advanced and could eventually escape human control.

OpenAI CEO Sam Altman has also recently supported proposals to slow down the development of the technology and introduce greater regulation.

While noting that these concerns may be justified, researchers have also questioned whether this is part of a diversion tactic to drum up investment and distract from the environmental damage AI data centers are currently causing.

Edited by: Elizabeth Schumacher

If you rely on our team for trusted reporting, please take a moment to select us as your Preferred Source on Google by clicking here and hitting the "star" or "preferred" button, so you'll always see our verified news first.

Skip next section DW's Top Story

DW's Top Story

Skip next section More stories from DW

More stories from DW