A new survey of 1,547 papers identifies the "horizon gap" in LLM agents, where models excel at short reasoning tasks but...
A new survey of 1,547 papers identifies the "horizon gap" in LLM agents, where models excel at short reasoning tasks but fail over multi-hour work due to lost context, premature co...