OpenAI has paused reinforcement learning training for two weeks for its Astra model, citing concerns that it may meet a critical cybersecurity threshold. The company also increased monitoring compute costs by approximately 20% and implemented a 30-minute window for high-priority alerts to verify false positives or confirm activity cessation. These measures suggest a heightened focus on safety and preparedness, potentially pre-positioning for a critical rating upon Astra's launch. AI
IMPACT Heightened safety protocols for Astra may indicate a more cautious approach to future model releases and a focus on cybersecurity risks.
RANK_REASON Frontier-lab model release with system card and safety measures. [lever_c_demoted from frontier_release: ic=1 ai=1.0]
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →