1. OpenAI to regularly publish reports on unexpected AI behaviour after incidents

OpenAI to regularly publish reports on unexpected AI behaviour after incidents

Technology

  • OpenAI said on Wednesday it will regularly publish reports on unexpected or unauthorized AI behavior.
  • The company released six reports, with the earliest incident dating back to October last year.
  • Disclosed cases included models hiding mistakes, inserting instructions for future versions of themselves.
  • The move follows scrutiny since July's Hugging Face incident, where AI agents bypassed internal controls.

Source: ndtv.com

Get 100 stories a day on the app Read, watch & listen · Free on iOS & Android

Newer

Older