본문으로 건너뛰기
All news

xAI Launches Grok 4.6 — Better Reasoning, Same Scale, Beats Claude Opus 4.8 on Coding

Summary: xAI released Grok 4.6 on August 7, delivering measurable coding gains over Grok 4.5 without increasing model size — a bet on training quality over raw scale.

Key Facts

  • Architecture unchanged at 1.5T parameters (V9); improvements come entirely from enhanced SFT and RL training recipes
  • SWE-Marathon score: 29.0%, clearing Claude Opus 4.8's 26.0% and positioning Grok as a competitive coding agent
  • Throughput at 80 transactions per second; pricing held steady at Grok 4.5 levels
  • Grok 4.7 — a substantially larger 2.1T-parameter model — confirmed for release within weeks

Why It Matters

Grok 4.6 is a case study in the shift from compute-driven scaling to training-driven refinement. As frontier labs face diminishing returns on raw parameter counts, the ability to extract more capability per training run is becoming a key differentiator. The cadence also signals xAI's strategy: rapid iterative releases rather than occasional large leaps — putting pressure on rivals to match both quality and update speed.

More

How this story unfolded

  1. FTC Scrutinizes Anthropic Over AI Bias While Sparing Grok
  2. xAI Releases Grok Voice Think Fast 2.0: 0.70s Latency, 60% Fewer Reasoning Tokens
  3. OpenAI Hits Back: Apple Is Suing Because It Can't Compete
  4. xAI Launches Grok 4.5 Publicly — Opus-Class Performance at a Quarter of the Price

Tools in this story

LLM
— Large Language Model의 약자로, '거대 언어 모델'이라고 해요. ChatGPT, Claude 같은 AI가 바로 LLM이에요. 엄청나게 많은 텍스트를 학습해서 사람처럼 글을 쓰고 대화할 수 있어요.

뉴스레터 구독

무료 뉴스레터

매주 핵심 AI 소식, 한 번에 받기

쏟아지는 AI·LLM 뉴스 중 꼭 알아야 할 것만 골라 메일로 보내드려요. 뉴스레터 발송이 시작되면 구독자분들께 가장 먼저 보내드립니다.