Research questionHow should vision-language models answer valid parts of compound queries while withholding unsafe or unanswerable parts?A single visual-language query can combine content that should be answered with content that should be corrected, withheld, or treated as unknown. All-or-nothing response behavior can therefore both provide unsafe answers and refuse valid components.