Results
Search-enabled answer quality across public benchmarks.
Cost
Candidate model cost and Search API cost per benchmark task and per search query.
Latency
Model time and search time per benchmark task and per search query.
Tokens
Candidate model input and output tokens per benchmark task, and quality against total token usage.
Searches
Search queries the candidate model runs per benchmark task.
Leaderboard Details
Sortable public Search API benchmark rows and per benchmark breakdown.
Example Tasks
Representative example tasks for each benchmark behind the Artificial Analysis Search Index.