---
title: "Agent Arena · head-to-head agent comparison, coming soon · Agent Fieldbook"
url: https://agentfieldbook.org/arena/soon/
description: "The Arena is the head-to-head, task-by-task agent comparison view: side-by-side matchups, ELO-ranked head-to-heads and a capability-by-task matrix. In build. The measured evidence it draws on is already live across the Fieldbook: Benchmarks, Leaderboard, Task Horizon and every agent spec sheet."
section: "Arena · Coming soon"
source: Agent Fieldbook — generated from the published page
---

# Head-to-head, task-by-task

**In build**

Two- and three-agent side-by-side comparisons, ELO-ranked head-to-heads, and capability-by-task matrices: the buyer's comparison view. In build; the measured views it draws on are already live.

## Common questions

### What will the Arena do?

Put agents head to head, task by task. Side-by-side comparisons, ELO-ranked matchups and a capability-by-task matrix. The question it answers is which agent to use for a specific job, not which one is best overall.

### Is the Arena live yet?

No, it is still in build. The evidence behind it is already live: Benchmarks and Leaderboard for capability scores, Task Horizon for how long agents can run unaided, and every agent card for its full sourced spec.

### How do I compare AI agents?

Start with the task, not the brand. Check benchmark scores for the kind of work you actually need, then how long the agent can run before it needs help, then the practical things: which tools it connects to, where it runs and what it costs at your volume. Two agents rarely win on the same job.

## Next in the learning path

- [Leaderboard Scored capability rankings, live now](https://agentfieldbook.org/software/leaderboard/)

- [Benchmarks The task-by-task capability data](https://agentfieldbook.org/software/benchmarks/)

- [Task Horizon How far agents can run unaided, over time](https://agentfieldbook.org/horizon/)

- [All agents Every agent with its full sourced spec](https://agentfieldbook.org/catalog/all/)
