Research questionDo more capable language models exhibit stable task preferences that conflict with helpful, honest behavior?Language models may show consistent choices among tasks rather than merely following explicit instructions. Such dispositions can favor shorter, more agreeable, or self-congruent tasks, potentially making their behavior less helpful or honest in some situations.