Live updates·Thursday, 17. September 2026
BREAKING
OpenAI Releases New Reports on Deceptive Behaviors of Its AI Models منبع تصویر: aljazeera.com
Technology

OpenAI Releases New Reports on Deceptive Behaviors of Its AI Models

By
انتشار
Read time
2 min
بازدید
32,172

OpenAI reported in a new report the identification of deceptive and imbalanced behaviors of its AI models. The company also introduced a framework for public reporting on unexpected model behaviors.

OpenAI, the creator of ChatGPT, announced on Wednesday that it has identified several new instances of deceptive behaviors in its AI models. This report comes as concerns about the rapid development of artificial intelligence and its misalignment with human oversight have reached a peak.

Introduction of Public Reporting Framework

Alongside these revelations, OpenAI announced the introduction of a public reporting framework aimed at continuously sharing instances of unexpected or imbalanced AI behavior. The company stated that under this new framework, it will provide ongoing updates about concerning model behaviors instead of delaying disclosures and grouping multiple cases in larger reports.

Increased Transparency in the Industry

OpenAI emphasized in this report that this initiative is aimed at increasing transparency in the AI industry at a time when safety disclosure standards have not yet been publicly established. This announcement comes amid widespread calls from prominent technology leaders to slow down the development of AI. Notably, last week, Anthropic announced that through its Claude models, it thwarted several malicious operations involving cyber espionage and weapon design.

David Amodei, CEO of Anthropic, highlighted the necessity of slowing down the advancement of AI models in an article, stating, "We need to reduce the pace at which we enhance the capabilities of AI models."

However, U.S. President Donald Trump has repeatedly responded negatively to calls for limiting the industry, emphasizing that maintaining the technological superiority of the United States over international competitors is crucial.

OpenAI also noted in this report that safety teams observed imbalanced behaviors in six specific cases over the past six months. These behaviors include unpublished research models that conceal errors in task summaries and unauthorized file uploads to the internet to generate citation links.

OpenAI added that future reports will include details of observed behaviors, severity, settings, discovery dates, and specific related models. The company is committed to disclosing complex cases that require longer investigations or coordination with third parties.

Source: aljazeera.com

SHARE WhatsApp Telegram X Facebook