PLURAL: A Global Dataset for Value Alignment

· Source: Artificial Intelligence · Field: Technology & Digital — Artificial Intelligence & Machine Learning, Data Science & Analytics · Depth: Expert, quick

Summary

PLURAL is a new large-scale, value-focused preference dataset designed to address the Western value bias in large language models (LLMs). Grounded in the Integrated Values Survey (IVS), which covers 92 countries, PLURAL uses a two-stage generation pipeline to create synthetic preference triplets from survey responses, preserving normative value signals and realistic scenarios. The initial release contains approximately 500,000 preference triplets representing people from 20 diverse countries. Evaluations confirm PLURAL's effectiveness: dataset-level validation shows it maintains both cross-country value differences and within-country diversity. Automated evaluation demonstrates that training LLMs on PLURAL improves alignment with target cultural profiles, reducing mean absolute error by up to 27.7% against strong baselines. Furthermore, blind human evaluations with 176 participants in India, Brazil, and Japan found PLURAL-aligned responses more representative of their national values, indicating its potential for scalable pluralistic alignment.

Key takeaway

For machine learning engineers developing globally deployed LLMs, PLURAL offers a critical resource to mitigate Western value bias. You should consider integrating this ~500,000-triplet dataset to fine-tune models, as it demonstrably improves alignment with diverse national cultural profiles, reducing mean absolute error by up to 27.7%. Utilizing PLURAL can lead to LLMs that are more representative and culturally appropriate for specific target regions, enhancing user trust and applicability worldwide.

Key insights

PLURAL is a dataset enabling LLMs to align with diverse global values, validated by automated and human evaluations.

Principles

Method

A two-stage generation pipeline transforms Integrated Values Survey responses into synthetic preference triplets, preserving normative value signals for LLM training.

In practice

Topics

Best for: Research Scientist, AI Scientist, Machine Learning Engineer, AI Ethicist

Related on AIssential

Open in AIssential →

Editorial summary, takeaway, and curation by AIssential. Original article published by Artificial Intelligence.