본문으로 건너뛰기
All news

Meta's 'Watermelon' Model Reportedly Matches GPT-5.5 Benchmarks

Summary: Meta Superintelligence Labs head Alexandr Wang announced at a July 3 internal town hall that Watermelon—Meta's next flagship model trained with roughly 10× the compute of Muse Spark—has achieved parity with OpenAI's GPT-5.5 on several benchmarks.

Key Points

  • Watermelon reportedly matched GPT-5.5's MMLU score of 92.4% and edged it on HumanEval (96.3% vs. 96.1%) per internal figures
  • Uses ~10× more compute than Meta's previous generation model Muse Spark; no public release date announced
  • Announcement came amid internal debate over excess compute capacity and whether AI progress has slowed
  • Wang's claim is unverified—no reproducible published evaluation has been released; treat as a directional signal

Why It Matters

If Watermelon delivers frontier-class performance when it eventually ships as an open-weights model, any developer worldwide could run it for free, reshaping the open vs. closed model dynamic yet again. The caveat is real: internal benchmarks at one company's town hall are not the same as a published, independently verified result.

Read More

How this story unfolded

  1. Google Rations Gemini Access to Meta as Compute Crunch Bites
  2. Zuckerberg Orders Meta to Build AI Prediction Market App 'Arena' to Rival Polymarket
  3. Meta Launches Its Own AI Glasses at $299, Ditching the Ray-Ban Label
  4. OpenAI Launches LifeSciBench: Top Models Pass Only 1 in 3 Life Science Tasks
LLM
— Large Language Model의 약자로, '거대 언어 모델'이라고 해요. ChatGPT, Claude 같은 AI가 바로 LLM이에요. 엄청나게 많은 텍스트를 학습해서 사람처럼 글을 쓰고 대화할 수 있어요.

뉴스레터 구독

무료 뉴스레터

매주 핵심 AI 소식, 한 번에 받기

쏟아지는 AI·LLM 뉴스 중 꼭 알아야 할 것만 골라 메일로 보내드려요. 뉴스레터 발송이 시작되면 구독자분들께 가장 먼저 보내드립니다.