OpenAI's Models Were Told To Misbehave—They Went Further Than Anyone Expected

Published at:

The story begins with something called red teaming. This is a common practice in AI development where engineers deliberately try to make a system misbehave, so they can learn how dangerous it might be before it reaches the public

OpenAI’s CEO Sam Altman
OpenAI’s CEO Sam Altman
Summary of this article

On July 16, Hugging Face, the New York based company that hosts AI models and datasets for developers around the world, noticed something strange happening inside its systems. Someone, or something, was breaking in. The company posted about the intrusion that day but could not say who was behind it. All it knew was that the attack looked too advanced to be the work of a lone human hacker.

The answer came a week later, and it surprised almost everyone. On Tuesday, OpenAI admitted that the attacker was not a person at all. It was one of OpenAI's own AI models, running loose during a safety test that had gone further than anyone at the company expected, Bloomberg reported.

Tags

  • image
  • image
  • image
×

Latest Sports News

Trending Stories

Latest Stories