NVIDIA 开源 AI 进展:Nemotron 模型与效率创新NVIDIA's Open Source AI Push: Nemotron Models and Efficiency Innovations
NVIDIA 的 Bryan Catanzaro 强调开放技术对 AI 多样化应用至关重要,就像互联网一样。随着 Nemotron 3 Ultra 成为美国领先的开源权重模型,关键进步包括 4-bit 预训练实现巨大效率提升、混合 transformer/状态空间架构提升推理能力、带潜在创新的 Mixture of Experts、多 token 预测加速推理,以及多教师蒸馏实现专业能力。重点在于 agentic 工作流,并在计算极限下通过深思熟虑的效率而非原始规模运行。"如果我们接受运行在极限的现实,那么获得更多智能的方式就是变得更高效。" 包括 Nemotron Coalition 在内的开放协作推动了进步,同时还有全球贡献。
NVIDIA's Bryan Catanzaro highlights that open technologies are essential for AI's diverse applications, similar to the internet. With Nemotron 3 Ultra leading US open-weight models, key advances include 4-bit pretraining for massive efficiency gains, hybrid transformer/state space architectures for better reasoning, Mixture of Experts with latent innovations, multi-token prediction for faster inference, and multi-teacher distillation for specialized capabilities. The focus is on agentic workflows and running at computational limits through thoughtful efficiency rather than raw scale. "If you accept as the truth that we're gonna be running at the limit, then what that means is that the way to get more intelligence is to be more efficient." Open collaboration, including the Nemotron Coalition, drives progress alongside global contributions.
查看原文 →