Live page · Day archive

ASCENT: Online Test-Time Training of Long-Horizon Agents via Self-Distillation of Verified Experience

Read the original at HF Daily Papers

Summary

Researchers introduce ASCENT, a method that trains large language model agents on their own verified execution trajectories during deployment to improve performance on related tasks.

Carried by: HF Daily Papers. First seen: .