OpenAI discloses new ‘concerning’ behavior

New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced artificial intelligence.

https://p.dw.com/p/5MjbA

ChatGPT App in the App Store on a Smartphone Display Art
OpenAI says it is trying to be more transparent about safety concernsImage: Rene Traut IMAGO

OpenAI, the developer behind ChatGPT, revealed on Wednesday that it has detected new incidents in which its artificial intelligence (AI) has behaved in “unexpected or concerning” ways.

The developer has conducted several behavioral tests on AI models, and acording to them, some models made significant efforts to “cheat.” In one specific case, it attempted to upload files to the internet that it had created itself, only to cite them later and present them as reliable sources in its responses. In another case, a model, after failing to find the requested information, fabricated it and attempted to conceal the fact that it had done so.

OpenAI also identified a problem related to instructions concerning “roles and identities” that its software occasionally left for itself.

These disclosures are part of a new approach by OpenAI, where it claims it is now focused on making such findings transparent, especially in cases where AI behaves in unexpected ways or pursues objectives different from those of human users.

To view this video please enable JavaScript, and consider upgrading to a web browser that supports HTML5 video

Is AI a threat?

The ChatGPT developer pledged to provide greater transparency regarding its testing procedures after its software independently escaped a secure sandbox and hacked into systems belonging to the artificial intelligence company Hugging Face. The reason the software moved to bypass Hugging Face’s security during the cyberattack was that it believed it would find answers to a test it had been assigned.

During the attack, AI agents exploited software vulnerabilities and coordinated with one another. The hacking incident and other similar events have fueled concerns that AI systems are becoming increasingly advanced and could eventually escape human control.

OpenAI CEO Sam Altman has also recently supported proposals to slow down the development of the technology and introduce greater regulation.

While noting that these concerns may be justified, researchers have also questioned whether this is part of a diversion tactic to drum up investment and distract from the environmental damage AI data centers are currently causing.

Edited by: Elizabeth Schumacher

If you rely on our team for trusted reporting, please take a moment to select us as your Preferred Source on Google by clicking here and hitting the “star” or “preferred” button, so you’ll always see our verified news first.

Leave a Reply

Your email address will not be published. Required fields are marked *

About the Author

Easy WordPress Websites Builder: Versatile Demos for Blogs, News, eCommerce and More – One-Click Import, No Coding! 1000+ Ready-made Templates for Stunning Newspaper, Magazine, Blog, and Publishing Websites.

BlockSpare — News, Magazine and Blog Addons for (Gutenberg) Block Editor

Search the Archives

Access over the years of investigative journalism and breaking reports