返回全部动态

新基准评估AI代理搜索API的质量、成本与速度

原标题:New benchmark ranks search APIs for AI agents on quality, cost, and speed

THE DECODER研究质量 72

AI 摘要

Artificial Analysis 发布了名为“Search Index”的新基准,用于评估搜索 API 提供商在 AI 代理场景中的质量、成本和速度。该基准使用统一模型 GPT-5.6 Luna 和开源框架 Stirrup 测试了 Parallel、Exa、Firecrawl 等七家提供商,并结合三个子基准进行评分。结果显示,更好的搜索质量能降低总成本,而原始速度并不总是带来更快的整体任务完成时间。

以上摘要由 AI 生成,可能存在误差。事实请以原文为准。

正文节选

New benchmark ranks search APIs for AI agents on quality, cost, and speed Artificial Analysis has released the "Search Index," a benchmark that measures how well search API providers work for AI agents across quality, cost, and speed. The initial lineup includes Parallel, Exa, Firecrawl, You.com, Tavily, Keenable, and Brave. Each one is tested with the same model (GPT-5.6 Luna) in a standardized agent setup. Only the search provider changes. The agent runs on Stirrup, an open-source framework fr


发布时间:2026-08-19 02:10
抓取时间:2026-08-19 03:02
来源机构:THE DECODER
阅读原文the-decoder.com