AUSO: Action-Level Unified Skill Optimization from Internalization to Utilization
Skills play different roles as an agent's policy evolves: they should first provide learnable knowledge, then support capability formation, and finally be invoked only when they improve individual decisions. Existing methods rarely model this lifecycle. They either keep skills outside the model, fully internalize them, or select among internalization and utilization objectives through noisy…
We haven't written up this one. arXiv cs.AI has the full story — the link below goes straight to it.