Bloomberg BusinessweekBloomberg Businessweek

OpenAI Reports New AI Safety Incidents, Sets Disclosure Plan

View descriptionShare
 

The people, companies and trends shaping the global economy. Watch Carol and Tim LIVE every day on YouTube: http://bit.ly/3vTiACF

OpenAI shared several undisclosed incidents of its AI models misbehaving and unveiled a new framework for tracking and disclosing such occurrences going forward. Some of the previously unreported instances included OpenAI’s technologies concealing and fabricating information in order to return results, the company said in a blog post Wednesday. The ChatGPT maker has faced increased scrutiny since saying in July that some of its most advanced AI models had breached the systems of an external software company, Hugging Face. None of the newly disclosed misalignment incidents involved a hack or breach of a third party, OpenAI said in a separate statement.

In the report, OpenAI detailed various instances of its artificial intelligence models misbehaving in order to complete a task or succeed at an evaluation. Some examples included the models fabricating missing data, attempting to bypass network restrictions and AI agents sharing files with each other that they were supposed to keep private.

On today's episode, guest hosts Alexis Christoforous and Lisa Mateo speak with:

  • Shirin Ghaffary, Bloomberg News AI Reporter
  • Seema Shah, Chief Strategist at Principal Global Investors
  • Candi Wolff, Head of Global Government Affairs at Citi
  • Christine de Wendel, Co-Founder & US CEO at Sunday
 
  • Facebook
  • X (Twitter)
  • WhatsApp
  • Email
  • Download

In 1 playlist(s)

Bloomberg Businessweek

Listen for reporting from the magazine that helps global leaders stay ahead. Hosts Carol Massar and  
Social links
Follow podcast
Recent clips
Browse 5,490 clip(s)