An AI agent designed and evaluated on tasks that require hundreds of dependent steps and tool calls sustained over extended time, rather than a single-turn or short-sequence response.
REAL-WORLD EXAMPLE
"It's not a chatbot demo anymore, it's a long-horizon agent running the whole migration unattended."
LORE & ORIGIN
Term of art from 2026 agent-benchmark literature, used across benchmarks like Long-Horizon-Terminal-Bench and SWE-Bench Pro and discussed in industry write-ups such as Arize AI's agent-evaluation field guide and arXiv 2607.08964. Distinguishes agents built and measured for sustained, multi-step autonomy from ones built only for quick single-shot completions.
MENTIONS OVER TIME
Radar observation data not yet available for this term.
RELATED SEARCHES
TOP PLATFORMS
Not enough Radar data yet to show a platform breakdown.
WHEN DID YOU FIRST HEAR THIS?
CLICK AN OPTION BELOW TO CAST YOUR VOTE.
[ VOTE TO REVEAL COMMUNITY RESULTS ]