Updated daily · Arena + Artificial Analysis
Agentic GTM Workflows
Signal monitoring, enrichment pipelines, multi-step outbound automations, CRM agents. The model plans, calls tools, and completes work end to end.
One board from The GTM Model Leaderboard
Agentic GTM Workflows Leaderboard
Signal monitoring, enrichment pipelines, multi-step outbound automations, CRM agents. The model plans, calls tools, and completes work end to end.
- 1Claude Fable 5 (High)Anthropic12.72%±2.00Win rate
- 2GPT-5.6 Sol (xHigh)OpenAI10.12%±1.69Win rate
- 3Claude Opus 4.8 (Thinking)Anthropic9.75%±1.39Win rate
- 4Kimi K3Moonshot AI9.71%±1.52Win rate
- 5Claude Sonnet 5 (High)Anthropic8.66%±1.89Win rate
Operator take · This is the category that decides whether your outbound agent finishes the job or stalls mid-pipeline. Claude models hold 3 of the top 5 agent slots; Sonnet 5 is the value pick when you run agents at volume.
Updated daily from the linked public leaderboards; last capture Jul 27, 2026. Elo and win-rate figures are preference-based measures, not task-completion guarantees. Always validate the top pick on your own representative work before routing production volume to it.
How to Read These Rankings
Four rules before you switch models
- Every ranking is task-specific. The #1 agent model is not the #1 writer, and neither is the right pick for 10,000 Clay rows.
- Scores come from public leaderboards (Arena, Artificial Analysis) with capture dates shown. Preference Elo measures which output people like, not whether work gets completed.
- Reasoning effort and harness matter: the same model at a different effort tier or in a different agent harness ranks differently.
- Cheap plus fast beats frontier for volume work. Route by job, not by brand loyalty.
We route these models inside production GTM systems every day. If you want help picking and wiring the right models into your own stack, that's what we do.