Research questionHow can hierarchical reinforcement learning use incrementally acquired knowledge for long-horizon exploration with sparse rewards?In fixed-knowledge HRL, information discovered during exploration does not readily alter the high-level structure used to choose subgoals. This makes long-horizon exploration inefficient when rewards are sparse and useful behavior must be learned from limited experience.