Code and lightweight result artifacts for Whose Norms? Disentangling Cultural and Personal Alignment in Large Language Models.
PACT evaluates whether language models choose to follow a cultural norm or allow a personal preference when the two are in tension.
scripts/ Core analysis and plotting scripts
data/model_behavior/ Cleaned benchmark/sample instances for model-behavior analysis
data/prompt_ablation/ Prompt-ablation release data
data/trace_analysis/ Trace-analysis tables and qualitative examples
figures/section4/ Generated Section 4 figures
figures/section5/ Generated Section 5 figures
scripts/compute_current_eval_metrics.py: compute model-evaluation summary metrics.scripts/plot_section4_pastel_model_behavior.py: generate Section 4 model-behavior figures.scripts/plot_section4_demographics_pastel.py: generate demographic/context plots.scripts/analyze_appendix_model_behavior.py: appendix analyses for ablations and preference types.scripts/section5_metrics.py: human-study and human-model alignment metrics.scripts/plot_section5_acl_figures.py: Section 5 figures.scripts/model_only_section5_analysis.py: model-only persona/no-persona analyses.
python3 -m venv .venv
source .venv/bin/activate
pip install -r requirements.txtThis release folder intentionally excludes raw scratch outputs, Slurm logs, API keys, and large intermediate model outputs. The included CSVs are lightweight summaries and cleaned release artifacts intended for paper reproduction and inspection.
@misc{borah2026pact,
title={Whose Norms? Disentangling Cultural and Personal Alignment in Large Language Models},
author={Borah, Angana and Augenstein, Isabelle and Mihalcea, Rada},
year={2026}
}