본문으로 건너뛰기
All news

NVIDIA Open-Sources Nemotron 3 Ultra 550B — Top US Open-Weights Model with 1M-Token Context and Commercial License

One-line summary: NVIDIA open-sourced Nemotron 3 Ultra, a 550B-parameter (55B active) hybrid Mamba-Transformer MoE model built for long-horizon agentic tasks — the strongest US open-weights release to date by benchmark intelligence score.

Key facts

  • Architecture: 550B total / 55B active parameters; interleaved Mamba-2, MoE, and selective Attention layers (LatentMoE design)
  • Context: 1 million tokens natively
  • Throughput: 300+ tokens/second
  • License: Linux Foundation OpenMDW-1.1 (commercially permissive)
  • Score: 48 on Artificial Analysis Intelligence Index — #1 among US open-weights, behind China's Kimi K2.6 at 54
  • Ships with 4 checkpoints (NVFP4, BF16 instruct, BF16 base, GenRM) plus training data and recipes
  • Available June 4 on Hugging Face, OpenRouter, and NVIDIA NIM

Why it matters

With Meta's Llama release cadence slowing, NVIDIA is staking a direct claim in the open-weights ecosystem — not just as chip supplier but as model provider. Nemotron 3 Ultra beats other US open models by a wide margin on agent workloads, but the gap with China's open frontier (Kimi K2.6, GLM-5) signals that US open-source still trails internationally, keeping the competitive pressure high.

Read more

How this story unfolded

  1. Alibaba's Qwen3-Coder-Next: 80B-Param MoE Coding Model With Only 3B Active — Scores 70.6% on SWE-Bench
  2. DeepSeek V4 Pro: 1.6T Open MoE Model Hits 80.6% SWE-bench at One-Tenth GPT-5.5's Price
  3. NVIDIA Launches Cosmos 3, the First Open Omnimodel for Physical AI
  4. NVIDIA Kicks Off Blackwell Ultra B300 Mass Production — 50x Hopper Throughput per Watt
LLM
— Large Language Model의 약자로, '거대 언어 모델'이라고 해요. ChatGPT, Claude 같은 AI가 바로 LLM이에요. 엄청나게 많은 텍스트를 학습해서 사람처럼 글을 쓰고 대화할 수 있어요.

뉴스레터 구독

무료 뉴스레터

매주 핵심 AI 소식, 한 번에 받기

쏟아지는 AI·LLM 뉴스 중 꼭 알아야 할 것만 골라 메일로 보내드려요. 뉴스레터 발송이 시작되면 구독자분들께 가장 먼저 보내드립니다.