返回全部动态

Ling-3.0-flash-VL 多模态模型 GGUF 量化版发布

原标题:bloomer010/Ling-3.0-flash-VL-GGUF

Hugging Face New and Trending Models一手来源开源质量 77

AI 摘要

Hugging Face 用户 bloomer010 发布了 inclusionAI/Ling-3.0-flash-VL 的 GGUF 量化版本,基于 llama.cpp 运行。该模型为 124B 总参数、每 token 仅激活 5.5B 的 MoE 多模态模型,支持图像和视频输入,原生 128K 上下文,可通过 YaRN 扩展至 256K。llama.cpp 已于 2026-09-24 合并支持(b11190 及更新版本),并提供多种量化规格及 DSpark 投机解码加速方案。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

--- license: mit base_model: - inclusionAI/Ling-3.0-flash-VL pipeline_tag: image-text-to-text library_name: llama.cpp tags: - gguf - bailingmoe3 - mixture-of-experts - vision - video - conversational --- # Ling-3.0-flash-VL GGUF GGUF conversions of [inclusionAI/Ling-3.0-flash-VL](https://huggingface.co/inclusionAI/Ling-3.0-flash-VL) > Built upon Ling-3.0-flash, it brings visual information into the complete process of understanding, reasoning, acting, and verification—advancing beyond image a


发布时间:2026-09-26 11:28
抓取时间:2026-09-26 11:16
来源机构:Hugging Face
阅读原文huggingface.co