The horizon gap: planning, memory, execution, training, and evaluation for long-horizon LLM agents
Read the original at arxiv.org→arXiv:2608.06663v1 Announce Type: new Abstract: Frontier language models solve reasoning problems in a single forward pass that would have been research contributions years ago, yet fail at multi-hour tasks: losing...
Original headline: "The Horizon Gap: Planning, Memory, Execution, Training, and Evaluation for Long-Horizon LLM Agents"
Coverage timeline
- Aug 10, 04:00 UTC arXiv cs.CL lead source The Horizon Gap: Planning, Memory, Execution, Training, and Evaluation for Long-Horizon LLM Agents