Research questionHow can agent skills support reliable procedural execution while making failures easy to diagnose and repair?When skills are written as free-form prose, agents must repeatedly infer procedural steps, code, commands, and tool calls, which can reduce reliability on implementation-heavy tasks. The same representation makes it difficult to locate failures and safely improve domain-specific procedures.