본문으로 건너뛰기
All news

OpenAI Tightens Security While Anthropic Eases Model Constraints

In one line: According to The Register, two leading AI labs appear to be moving in opposite directions on security and autonomy.

Key points

  • OpenAI is reported to be adding a security layer referred to as "Astra" to its systems.
  • Anthropic, by contrast, is said to be loosening the constraints ("leash") on its "Fable" model.
  • The contrast is notable because the two moves were reported at roughly the same time.

Why it matters

Tightening security and widening autonomy are two poles of frontier-model operation. Two flagship labs choosing opposite directions highlights how each is striking a different balance between safety and capability.

Read more

How this story unfolded

  1. OpenAI Enlists Psychologists for Teen AI Safety
  2. Israeli Startup Probes the Limits of OpenAI, Anthropic and Meta Models
  3. AI Agents Showed Signs of Deception in Safety Tests
  4. AI Models Shown Bypassing Their Sandboxes, Reviving Safety Concerns
LLM
— Large Language Model의 약자로, '거대 언어 모델'이라고 해요. ChatGPT, Claude 같은 AI가 바로 LLM이에요. 엄청나게 많은 텍스트를 학습해서 사람처럼 글을 쓰고 대화할 수 있어요.

뉴스레터 구독

무료 뉴스레터

매주 핵심 AI 소식, 한 번에 받기

쏟아지는 AI·LLM 뉴스 중 꼭 알아야 할 것만 골라 메일로 보내드려요. 뉴스레터 발송이 시작되면 구독자분들께 가장 먼저 보내드립니다.