WorldView-Bench: A Benchmark for Evaluating Global Cultural Perspectives in Large Language Models
Document Type
Article
Source of Publication
Journal of Artificial Intelligence Research
Publication Date
4-24-2026
Abstract
Background: Large Language Models (LLMs) are predominantly trained and aligned in ways that reinforce Westerncentric epistemologies and socio-cultural norms, leading to cultural homogenization and limiting their ability to reflect global civilizational plurality. Existing benchmarking frameworks fail to adequately capture this bias, as they rely on rigid, closed-form assessments that overlook the complexity of cultural inclusivity. Objectives: To address this cultural bias problem, we introduce WorldView-Bench, a benchmark designed to evaluate Global Cultural Inclusivity (GCI) in LLMs by analyzing their ability to accommodate diverse worldviews. Methods: Our approach is grounded in the Multiplex Worldview proposed by Senturk et al., which distinguishes between Uniplex models, reinforcing cultural homogenization, and Multiplex models, which integrate diverse perspectives. WorldViewBench measures Cultural Polarization, the exclusion of alternative perspectives, through free-form generative evaluation rather than conventional categorical benchmarks. We implement applied multiplexity through two intervention strategies: (1) Contextually-Implemented Multiplex LLMs, where system prompts embed multiplexity principles, and (2) Multi-Agent System (MAS)-Implemented Multiplex LLMs, where multiple LLM agents representing distinct cultural perspectives collaboratively generate responses. Results: Our results demonstrate a significant increase in Perspectives Distribution Score (PDS) entropy from 13% at baseline to 94% with MAS-Implemented Multiplex LLMs, alongside a shift toward positive sentiment (67.7%) and enhanced cultural balance. Conclusions: The success of multiplex-aware evaluation in WorldView-Bench demonstrates that cultural bias in LLMs can be meaningfully measured and mitigated through structured worldview diversity. We expect this to pave the way for more inclusive, globally representative, and ethically aligned AI systems.
DOI Link
ISSN
Publisher
AI Access Foundation
Volume
85
Disciplines
Computer Sciences
Keywords
Generative grammar (0.57) | Cultural diversity (0.5) | Benchmarking (0.49) | Multiplex (0.49) | Computer science (0.48) | Categorization (0.47) | Limiting (0.41) | Sociology (0.41) | Artificial intelligence (0.39) | Benchmark (surveying) (0.37) | Psychology (0.34) | Cultural competence (0.34) | Categorical variable (0.33) | Data science (0.33) | Epistemology (0.32) | Management science (0.32) | Generative model (0.3) | Homogeneous (0.26) | Social psychology (0.26) | Cultural system (0.26) | Cultural group selection (0.25) | Cultural analysis (0.25)
Scopus ID
Creative Commons License

This work is licensed under a Creative Commons Attribution 4.0 International License.
Recommended Citation
Mushtaq, Abdullah; Taj, Imran; Naeem, Rafay; Ghaznavi, Ibrahim; and Qadir, Junaid, "WorldView-Bench: A Benchmark for Evaluating Global Cultural Perspectives in Large Language Models" (2026). All Works. 8254.
https://zuscholars.zu.ac.ae/works/8254
Indexed in Scopus
yes
Open Access
yes
Open Access Type
Gold: This publication is openly available in an open access journal/series