Research questionHow can weight pruning make LLMs more efficient without amplifying prompt-dependent demographic bias?Pruning removes weights to improve LLM efficiency, but it can make outputs more sensitive to persona cues and amplify demographic bias. Deployment therefore requires balancing sparsity with model quality and fairness.