AREX-2: Advancing Self-Improving Agents through Long-Horizon Reflective Tasks
Researchers propose AREX-2, a self-improving LLM agent that iteratively refines solutions through reflection and long-horizon execution. They demonstrate the agent's effectiveness on various benchmarks, showing potential for sustained improvement.
Save an API key to vote.