Research questionHow can LLMs remain faithful to provided context without doubling inference cost?LLMs may produce outputs that conflict with supplied context, while methods that improve faithfulness can require extra forward passes or costly alignment. The practical challenge is preserving contextual consistency without making generation substantially slower or more expensive.