返回全部动态
Hugging Face 发布 Luth-2 法语小模型,采用 MOPD 技术刷新多项基准
原标题:Luth-2: Pushing the French Capabilities of SLMs with MOPD
AI 摘要
Hugging Face 发布了 Luth-2-0.8B 和 Luth-2-2B 两个法语小语言模型,基于 Qwen3.5 后训练,采用多域策略蒸馏(MOPD)技术,在数学、代码、指令跟随、通用知识和工具调用等基准上刷新了同尺寸模型的最优成绩,性能可与两到三倍大小的模型竞争。模型、数据、代码和评估方法均已开源。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
We introduce Luth-2-0.8B and Luth-2-2B, two models that set a new state of the art in French for their size across math, code, instruction following, general knowledge and tool calling. They are non-reasoning models post-trained from the Qwen3.5 family in French. Luth-2 builds on our previous work, Luth, with several substantial improvements. Our new SFT mixture, Luth-2-Post-Training-SFT, has grown to 3B tokens and covers a broader range of domains, including mathematics, knowledge, code, tool c
发布时间:—
抓取时间:2026-08-13 16:54
来源机构:Hugging Face