Research questionHow can we reliably detect when an LLM response is unsupported by its reference documents?LLM outputs can contain factual claims absent from or contradicted by their source documents, while opaque generation provides little evidence for checking them. Detection must assess whether responses are grounded in those references without domain-specific fine-tuning.