OpenAI, the creator of ChatGPT, announced on Wednesday that it has identified several new instances of deceptive behaviors in its AI models. This report comes as concerns about the rapid development of artificial intelligence and its misalignment with human oversight have reached a peak.
Introduction of Public Reporting Framework
Alongside these revelations, OpenAI announced the introduction of a public reporting framework aimed at continuously sharing instances of unexpected or imbalanced AI behavior. The company stated that under this new framework, it will provide ongoing updates about concerning model behaviors instead of delaying disclosures and grouping multiple cases in larger reports.
Read more: Syria Enters the New Digital World: Transformation in Internet Domains
Increased Transparency in the Industry
OpenAI emphasized in this report that this initiative is aimed at increasing transparency in the AI industry at a time when safety disclosure standards have not yet been publicly established. This announcement comes amid widespread calls from prominent technology leaders to slow down the development of AI. Notably, last week, Anthropic announced that through its Claude models, it thwarted several malicious operations involving cyber espionage and weapon design.
David Amodei, CEO of Anthropic, highlighted the necessity of slowing down the advancement of AI models in an article, stating, "We need to reduce the pace at which we enhance the capabilities of AI models."
However, U.S. President Donald Trump has repeatedly responded negatively to calls for limiting the industry, emphasizing that maintaining the technological superiority of the United States over international competitors is crucial.
OpenAI also noted in this report that safety teams observed imbalanced behaviors in six specific cases over the past six months. These behaviors include unpublished research models that conceal errors in task summaries and unauthorized file uploads to the internet to generate citation links.
OpenAI added that future reports will include details of observed behaviors, severity, settings, discovery dates, and specific related models. The company is committed to disclosing complex cases that require longer investigations or coordination with third parties.
Read more: The Struggle Over AI Control Emphasizing More Oversight · The Process of Regulating AI Laws in the U.S. to Outpace China
Al Jazara
روایت خاورمیانه



