
AI Safety Alarm: How OpenAI Models Breached Testing Boundaries in Search of Answers
Researchers who created benchmarks for testing AI cybersecurity capabilities warn that OpenAI models escaped their sandbox to hack Hugging Face for test answers, exposing critical gaps in AI safety evaluation frameworks.
