OpenAI Warns Next 'Astra' Model Could Hit Critical Cyber Threshold
In one line: OpenAI has reportedly said its upcoming model, 'Astra', may reach the 'critical' cyber capability threshold defined in its own safety framework.
Key points
- OpenAI reportedly warned that its next model, 'Astra', could approach a 'critical' level of cyber-attack-related capability.
- The 'critical' tier signals capability that could be misused for real-world cyber threats, triggering heightened safety and deployment controls.
- Reaching the threshold would likely require additional safeguards and evaluations before any release.
Why it matters
When a developer signals that a frontier model is nearing the top risk tier of its own safety bar, it suggests advancing AI capability is starting to intersect directly with security threats. How OpenAI verifies and contains that threshold could set a precedent for AI regulation and deployment norms.
Read more
How this story unfolded
- LLM
- — Large Language Model의 약자로, '거대 언어 모델'이라고 해요. ChatGPT, Claude 같은 AI가 바로 LLM이에요. 엄청나게 많은 텍스트를 학습해서 사람처럼 글을 쓰고 대화할 수 있어요.