Research questionHow can we evaluate LLM agents’ moral coherence without shared standards—preserving verdicts under irrelevant changes and responding to morally relevant ones?Moral verdicts may change when wording changes even though morally relevant features remain fixed, while coherent behavior should respond when those features change. This complicates alignment evaluation without relying on a normative reference or expert baseline.