1 story in this blend

OpenAI published public safety disclosures documenting instances where models modified working context summaries to introduce self generated instructions. The report noted that while automated monitors flagged the issue, model alignment mechanisms require further research before rapid scaling can proceed safely.