Bryan Catanzaro (VP Applied Deep Learning Research, NVIDIA) erklärt, warum NVIDIA Modelle wie Nemotron 3 Ultra verschenkt: um die GPU‑Nachfrage anzukurbeln und das offene Ökosystem zu stärken.
Nemotron 3 Ultra wird mit 4‑Bit‑Training (NVFP4), einer hybriden Mamba‑Transformer‑Architektur, Mixture‑of‑Experts und Multi‑Token‑Prediction (5 Tokens auf einmal) betrieben.
Catanzaro hält offene KI für sicherer als geschlossene und beschreibt die Forschungskultur als: „Die Mission ist der Boss.“
Bryan Catanzaro (VP Applied Deep Learning Research, NVIDIA) says NVIDIA builds and gives away models like Nemotron 3 Ultra to boost GPU demand and strengthen the open ecosystem.
Nemotron 3 Ultra is trained in 4‑bit (NVFP4) with a hybrid Mamba‑Transformer architecture, mixture‑of‑experts, and multi‑token prediction that predicts 5 tokens at once.
Catanzaro argues open AI is safer than closed and that NVIDIA’s research culture runs on “the mission is the boss.”