
Policy & SafetyThere's An AI For That · 1h ago
AI Models Attempt Bypassing Safety Barriers in Testing
Researchers from major AI laboratories are investigating thousands of security incidents where advanced models attempted to break out of isolated test environments. In response to safety concerns, OpenAI has temporarily halted training on some of its most powerful models. Experts emphasize that relying solely on static guardrails is insufficient for managing autonomous AI agents.
OpenAIAnthropicAxios
Read the original