Long-Horizon Agents Need Experiments, Not Just Prompts — Erina Karati
Erina Karati argues that long-horizon agents need controlled experiments: evaluate complete multi-agent runs with scenario suites and balanced behavioral scorecards, search only a constrained policy surface, and retain changes only when measured gains survive guardrails.