#long-horizon tasks
Long-Horizon Tasks: 4 AI articles covering long-horizon tasks news, analysis, and research
Articles
Context Engineering Inside the Harness: 4 Mechanisms That Beatβ7
Long-horizon AI agents fail from context overflow and goal loss. Learn how compaction, memory, budgeting, and todo-state in the harness fix it.
Subagents vs. Agent Skills: How to Execute Reusable Knowledgeβ8
Subagent execution outperforms loading agent skills into context for long-horizon tasks. New arXiv research shows how organizing reusable knowledge as subagents...
Nvidia Just Showed the Harness, Not the AI Model, Is Now the Real Heroβ9
New Nvidia research reveals the AI harnessβnot the modelβis the real hero for long-horizon tasks, boosting scores from 30% to 100% on ARC-AGI-3.
Z.ai Ships GLM-5.3 Without Retraining the Base Model: Better atβ8
Z.ai's GLM-5.3 boosts coding and cybersecurity via post-training, no retraining. Scores surge on complex benchmarks; open weights release in two weeks.
