返回全部动态

Nvidia发布Nemotron 3.5 Lightning:小参数高速度,性能对标GPT-oss-120b

原标题:Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence

THE DECODER模型发布质量 77

AI 摘要

Nvidia发布了开源权重模型Nemotron 3.5 Lightning,该模型拥有316亿总参数但仅激活36亿参数,在智能指数上达到24分,与OpenAI的gpt-oss-120b持平,但推理速度接近每秒670个token,远超同类模型。该模型在代理基准测试中表现显著提升,采用宽松的OpenMDW-1.1许可证,并提供BF16和NVFP4权重,多家云服务商已提供无服务器推理。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence Nvidia's new Nemotron 3.5 Lightning is a compact open-weights model that matches OpenAI's gpt-oss-120b on intelligence benchmarks with a quarter of the parameters while delivering the fastest inference speeds in its class. Nvidia has released Nemotron 3.5 Lightning, the first model in its new Nemotron 3.5 lineup. The model directly succeeds the Nemotron 3 Nano 30B A3B and keeps its hybrid Mamba-Transformer ar


发布时间:2026-08-11 23:07
抓取时间:2026-08-12 01:51
来源机构:THE DECODER
阅读原文the-decoder.com