返回全部动态
OpenCompass v0.5.2 发布:新增多项基准测试与模型支持
原标题:0.5.2
AI 摘要
OpenCompass 发布 v0.5.2 版本,新增多个科学和通用基准测试支持,包括 SciReasoner、Biology Instructions、Mol Instructions、CMPhysBench、IFBench、LCB-pro 等。同时增加了对 Intern-S1-Pro 模型和 TeleChat API 的评估支持,并修复了多个 bug,改进了评估流程和 CI/CD。
以上摘要由 AI 生成,可能存在误差。事实请以原文为准。
正文节选
# OpenCompass v0.5.2 Release Notes ## 🌟 **Highlights** ✨ **🧪 Extensive New Benchmarks Support**: We have introduced comprehensive support for Scientific and General Benchmarks, including **SciReasoner**, **Biology Instructions**, **Mol Instructions**, **CMPhysBench**, **IFBench**, **LCB-pro**, etc. ✨ **🤖 New Model & API Support**: Added support for **Intern-S1-Pro** and **TeleChat API** evaluation examples. ✨ **🛠️ Infrastructure & Enhancements**: Fixed bugs, improved evaluation pipelin
发布时间:2026-02-14 11:46
抓取时间:2026-08-31 00:41
来源机构:Shanghai AI Laboratory