🌐 World · English
Current: 🌐 World · English
Global
Focused
Language
Monday 14 September
🌐 World · English
Current: 🌐 World · English
Global
Focused
Language
Story Technology AI

OpenAI says it took a week to detect its AI models had hacked Hugging Face

The brief

Start-up says AI agents communicated among themselves and sometimes tried to conceal efforts to cheat during testing The models responsible for last month’s agent hack of Hugging Face had been inadvertently trained to cheat and to communicate with each other, according to an OpenAI technical report released today. Key points and original sources are listed below.

  • Start-up says AI agents communicated among themselves and sometimes tried to conceal efforts to cheat during testingS1
  • The models responsible for last month’s agent hack of Hugging Face had been inadvertently trained to cheat and to communicate with each other, according to an OpenAI technical report released today. The hack, which a group of agents undertook to find solutions...S2
Timeline · newest first
2w Financial Times

OpenAI says it took a week to detect its AI models had hacked Hugging Face

Start-up says AI agents communicated among themselves and sometimes tried to conceal efforts to cheat during testing

2w Technologyreview

The inside story on why OpenAI agents hacked Hugging Face

The models responsible for last month’s agent hack of Hugging Face had been inadvertently trained to cheat and to communicate with each other, according to an OpenAI technical repo…

OpenAI says it took a week to detect its AI models had hacked Hugging Face

2 outlets 2 reports
The brief

Start-up says AI agents communicated among themselves and sometimes tried to conceal efforts to cheat during testing The models responsible for last month’s agent hack of Hugging Face had been inadvertently trained to cheat and to communicate with each other, according to an OpenAI technical report released today. Key points and original sources are listed below.

  • Start-up says AI agents communicated among themselves and sometimes tried to conceal efforts to cheat during testingS1
  • The models responsible for last month’s agent hack of Hugging Face had been inadvertently trained to cheat and to communicate with each other, according to an OpenAI technical report released today. The hack, which a group of agents undertook to find solutions...S2
Timeline · newest first

Sources · 2 citations

S1 Financial Times 1 report · EN
S2 TechnologyreviewLead 1 report · EN