arXiv · 2503.16094
Cultural Alignment in Large Language Models Using Soft Prompt Tuning
Abstract
Large Language Model (LLM) alignment is commonly achieved through supervised fine-tuning or reinforcement learning, both of which require labeled or preference data and update model weights. Without targeted cultural adaptation, however, deployed LLMs often exhibit culturally homogeneous behavior that fails to reflect diverse local values. Aligning models to cultural value frameworks such as Hofstede's Value Survey Module (VSM13) presents a distinct challenge: alignment signals are available only as aggregated survey-level scores computed after generating responses to an entire survey, providing no per-token gradient and requiring no preference data by construction. This makes standard gradient-based alignment methods ill-suited to the task. We propose a deployment-friendly approach that encodes cultural behavior in short, tunable soft prompts optimized with Differential Evolution (DE), while keeping model weights frozen and requiring no preference data. At inference, the system inserts the appropriate cultural-specific prompt to adapt model responses for different cultures. Experiments across four countries and four instruction-tuned models show that DE-optimized prompts generally reduce discrepancy with VSM13 reference profiles, improve rank agreement with the World Values Survey (WVS), an independent framework not seen during optimization, and are preferred in blinded pairwise evaluations using majority voting across three LLM judges.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Reem I. Masoud, Martin Ferianc, Philip Treleaven, Miguel Rodrigues. 2026-09-18. Cultural Alignment in Large Language Models Using Soft Prompt Tuning. https://arxiv.org/abs/2503.16094
Cite the original work for its findings. Save a collection to share your selection of sources.