新基准评估AI代理搜索API的质量、成本与速度
原标题:New benchmark ranks search APIs for AI agents on quality, cost, and speed
AI 摘要
Artificial Analysis 发布了名为“Search Index”的新基准,用于评估搜索 API 提供商在 AI 代理场景中的质量、成本和速度。该基准使用统一模型 GPT-5.6 Luna 和开源框架 Stirrup 测试了 Parallel、Exa、Firecrawl 等七家提供商,并结合三个子基准进行评分。结果显示,更好的搜索质量能降低总成本,而原始速度并不总是带来更快的整体任务完成时间。
正文节选
New benchmark ranks search APIs for AI agents on quality, cost, and speed Artificial Analysis has released the "Search Index," a benchmark that measures how well search API providers work for AI agents across quality, cost, and speed. The initial lineup includes Parallel, Exa, Firecrawl, You.com, Tavily, Keenable, and Brave. Each one is tested with the same model (GPT-5.6 Luna) in a standardized agent setup. Only the search provider changes. The agent runs on Stirrup, an open-source framework fr