NVIDIA 发布 Nemotron-3-Super 120B 混合架构大模型
原标题:nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8
AI 摘要
NVIDIA 发布了 Nemotron-3-Super-120B-A12B-FP8 大型语言模型,采用 LatentMoE 混合架构,总参数 120B,激活参数 12B,支持最长 1M token 的上下文,并内置多 token 预测(MTP)以加速生成。该模型针对智能体工作流、长上下文推理和高吞吐量场景(如 IT 工单自动化)进行了优化,支持多种语言,并采用 NVIDIA 开放模型许可。模型已可在 Hugging Face 上下载,并提供了技术报告和预训练/后训练数据集。
正文节选
--- library_name: transformers license: other license_name: nvidia-nemotron-open-model-license license_link: >- https://www.nvidia.com/en-us/agreements/enterprise-software/nvidia-nemotron-open-model-license/ pipeline_tag: text-generation language: - en - fr - es - it - de - ja - zh tags: - nvidia - pytorch - nemotron-3 - latent-moe - mtp datasets: - nvidia/nemotron-post-training-v3 - nvidia/nemotron-pre-training-datasets track_downloads: true --- # NVIDIA-Nemotron-3-Super-120B-A12B-FP8 <div