Story
arxiv_cs_ai ยท Sep 10, 2026 ยท paper
arxiv.orgSep 10, 2026
original source linked
In brief
Large language model (LLM) agents can benefit from reusable skills distilled from prior task experience, yet existing skill optimization methods often rely on costly execution-based evaluation and substantial task dat...
Feed lens
agentharnessevaluation