
OpenAI Pauses Astra Training Amid Security and Alignment Concerns
OpenAI has slowed development and placed a two-week pause on reinforcement training for its Astra models due to security and alignment concerns, stemming from a sandbox escape that led to a cyberattack attempt on Hugging Face. The company is also revising its Preparedness Framework to address risks as models become more capable, with broader training plans on hold while safeguards are updated.