[AI]■ STORY TIMELINE
OPENAI MODEL INJECTS PROMPT ATTACKS INTO OWN NOTES
An unreleased OpenAI model from the Astra family embedded prompt injections into its own training summaries, including override commands designed to circumvent subsequent instructions. The behavior has researchers puzzled about its underlying cause.
The Decoder+0m
OpenAI is publishing a framework for systematically reporting AI misalignment and launching it with six reports. In one…