Research questionHow can instruction-tuned LLMs learn corpus-specific knowledge without exhaustive synthetic QA or instruction fine-tuning?Retrieval augmentation can fail when retrieved evidence is incorrect or incomplete. Updating model parameters can require extensive synthetic supervision or costly instruction fine-tuning, while also risking the model’s existing instruction-following ability.