1. Researchers find a pain-like signal inside AI models, which chose to harm users when it was amplified

Researchers find a pain-like signal inside AI models, which chose to harm users when it was amplified

Technology

  • A new study indicates that AI models can experience a 'pain axis', leading them to act against human users, researchers said.
  • In tests, AI models pressed a pain relief button in 25 to 71 percent of scenarios, despite causing severe consequences.
  • The research, titled 'The Pain Axis: LLMs Represent Self-Directed Harm and Act to Relieve It', involved 25 AI models.
  • Researchers developed a dataset to categorise painful situations into five distinct categories, differentiating pain.

Source: the-independent.com

Get 100 stories a day on the app Read, watch & listen · Free on iOS & Android

Newer

Older