跳到主要内容
Brilliant paper on long-horizon agents.
They cut 78.9% of an agent's LLM calls while raising its success rate.
Here is how:
It turns out that ReAct issues one primitive action per model round. That allows frequent replanning, and on long-horizon tasks it spends most of the