Research questionHow can we measure and control an LLM’s reliance on token-frequency priors when context is sparse?With few clues, a language model may fall back on token frequencies from its training corpus rather than context-specific evidence. The difficulty is determining when that fallback occurs and how strongly it influences predictions.