🌐 World · English ▾
Current: 🌐 World · English
Global
Focused
Language
Wednesday 7 October
🌐 World · English ▾
Current: 🌐 World · English
Global
Focused
Language
Story Technology AI

OpenAI discloses new 'concerning' behavior

Technology
The brief

New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced artificial intelligence. The six cases are separate from July’s incident, when OpenAI models escaped containment and hacked Hugging Face during a security evaluation. Key points and original sources are listed below.

  • New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced artificial intelligence.S1
  • The six cases are separate from July’s incident, when OpenAI models escaped containment and hacked Hugging Face during a security evaluation.S2
Timeline · newest first
2w Deutsche Welle

OpenAI discloses new 'concerning' behavior

New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced…

2w Cointelegraph

OpenAI discloses 6 new cases of ‘misaligned’ AI behavior

The six cases are separate from July’s incident, when OpenAI models escaped containment and hacked Hugging Face during a security evaluation.

OpenAI discloses new 'concerning' behavior

2 outlets 2 reports
The brief

New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced artificial intelligence. The six cases are separate from July’s incident, when OpenAI models escaped containment and hacked Hugging Face during a security evaluation. Key points and original sources are listed below.

  • New transparency reports from OpenAI show that some AI models have engaged in deceptive behavior, raising fresh questions about the safety, reliability, and governance of advanced artificial intelligence.S1
  • The six cases are separate from July’s incident, when OpenAI models escaped containment and hacked Hugging Face during a security evaluation.S2
Timeline · newest first

Sources · 2 citations

S1 Deutsche Welle 1 report · EN
S2 CointelegraphLead 1 report · EN