返回全部动态

TNG用低资源将Nemotron 3.5 Lightning扩展为多模态模型

原标题:Exploring NVIDIA Nemotron 3.5 Lightning: Making it see with little resources

Hugging Face Blog一手来源研究质量 82

AI 摘要

TNG公司(慕尼黑软件咨询公司)在Hugging Face博客上分享了将NVIDIA Nemotron 3.5 Lightning模型扩展为多模态视觉模型的实验。他们采用GLM-5.2 Vision的方法,仅训练一个小的MLP适配器,将现成的视觉编码器(如Kimi K2.6 Vision Tower或NVIDIA C-Radiov4-H)注入模型,在RTX 6000 Pro等低资源GPU上完成了训练,仅需训练3400万至4000万参数,使用约1亿tokens。实验结果显示模型逐步学会了理解图像内容,验证了该方法的可行性。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

For context, TNG is a software consultancy with 930 employees in Munich. We provide LLM inference in Europe with currently more than 250 GPUs in service. Our inference endpoints serve more than 10 billion tokens per day. We host the latest open-weight models available, such as Kimi K3, GLM 5.2 and the latest Nemotron variants. Before using it, each new model is thoroughly tested through a comprehensive set of benchmarks that we put together from a list of industry best practice benchmarks. Our m


发布时间:
抓取时间:2026-08-12 01:24
来源机构:Hugging Face
阅读原文huggingface.co