본문으로 건너뛰기
All news

OpenAI Warns Next 'Astra' Model Could Hit Critical Cyber Threshold

In one line: OpenAI has reportedly said its upcoming model, 'Astra', may reach the 'critical' cyber capability threshold defined in its own safety framework.

Key points

  • OpenAI reportedly warned that its next model, 'Astra', could approach a 'critical' level of cyber-attack-related capability.
  • The 'critical' tier signals capability that could be misused for real-world cyber threats, triggering heightened safety and deployment controls.
  • Reaching the threshold would likely require additional safeguards and evaluations before any release.

Why it matters

When a developer signals that a frontier model is nearing the top risk tier of its own safety bar, it suggests advancing AI capability is starting to intersect directly with security threats. How OpenAI verifies and contains that threshold could set a precedent for AI regulation and deployment norms.

Read more

How this story unfolded

  1. AI Agent Incidents Raise a New Enterprise Security Question
  2. OpenAI Missed Its AI Agents Coordinating a Hack Via a Message Board
  3. Anthropic AI Safety Test Reportedly Generated Fake Online Identities
  4. When AI Attacks AI: The Rise of Bot-vs-Bot Security Threats
LLM
— Large Language Model의 약자로, '거대 언어 모델'이라고 해요. ChatGPT, Claude 같은 AI가 바로 LLM이에요. 엄청나게 많은 텍스트를 학습해서 사람처럼 글을 쓰고 대화할 수 있어요.

뉴스레터 구독

무료 뉴스레터

매주 핵심 AI 소식, 한 번에 받기

쏟아지는 AI·LLM 뉴스 중 꼭 알아야 할 것만 골라 메일로 보내드려요. 뉴스레터 발송이 시작되면 구독자분들께 가장 먼저 보내드립니다.