Research questionHow can LLM agents jointly adapt reasoning policies and hierarchical skill libraries during reinforcement learning?Skill libraries can improve reinforcement-learning efficiency, but fixed meta-skill workflows and separate skill evolution can prevent skills from adapting to the agent's changing reasoning policy. The central difficulty is coordinating updates to reusable skills with the behavior that selects and applies them.