WorldView-Bench: A Benchmark for Evaluating Global Cultural Perspectives in Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Mushtaq, Abdullah, Taj, Imran, Naeem, Rafay, Ghaznavi, Ibrahim, Qadir, Junaid |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Toward Inclusive Educational AI: Auditing Frontier LLMs through a Multiplexity Lens
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
Harnessing Multi-Agent LLMs for Complex Engineering Problem-Solving: A Framework for Senior Design Projects
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
Can LLMs Write Faithfully? An Agent-Based Evaluation of LLM-generated Islamic Content
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
Can Agents Judge Systematic Reviews Like Humans? Evaluating SLRs with LLM-based Multi-Agent System
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025)
IslamicLegalBench: Evaluating LLMs Knowledge and Reasoning of Islamic Law Across 1,200 Years of Islamic Pluralist Legal Traditions
di: Elmahjub, Ezieddin, et al.
Pubblicazione: (2026)
di: Elmahjub, Ezieddin, et al.
Pubblicazione: (2026)
PolicySimEval: A Benchmark for Evaluating Policy Outcomes through Agent-Based Simulation
di: Kang, Jiaju, et al.
Pubblicazione: (2025)
di: Kang, Jiaju, et al.
Pubblicazione: (2025)
Using Large Language Models to Simulate Human Behavioural Experiments: Port of Mars
di: Slumbers, Oliver, et al.
Pubblicazione: (2025)
di: Slumbers, Oliver, et al.
Pubblicazione: (2025)
MultiAgentBench: Evaluating the Collaboration and Competition of LLM agents
di: Zhu, Kunlun, et al.
Pubblicazione: (2025)
di: Zhu, Kunlun, et al.
Pubblicazione: (2025)
Reimagining Urban Science: Scaling Causal Inference with Large Language Models
di: Xia, Yutong, et al.
Pubblicazione: (2025)
di: Xia, Yutong, et al.
Pubblicazione: (2025)
Research on Comprehensive Classroom Evaluation System Based on Multiple AI Models
di: Xie, Cong, et al.
Pubblicazione: (2025)
di: Xie, Cong, et al.
Pubblicazione: (2025)
Origin-Destination Pattern Effects on Large-Scale Mixed Traffic Control via Multi-Agent Reinforcement Learning
di: Fan, Muyang, et al.
Pubblicazione: (2025)
di: Fan, Muyang, et al.
Pubblicazione: (2025)
Enhancing Collective Intelligence in Large Language Models Through Emotional Integration
di: Kadiyala, Likith, et al.
Pubblicazione: (2025)
di: Kadiyala, Likith, et al.
Pubblicazione: (2025)
CompARE: A Computational framework for Airborne Respiratory disease Evaluation integrating flow physics and human behavior
di: Leong, Fong Yew, et al.
Pubblicazione: (2025)
di: Leong, Fong Yew, et al.
Pubblicazione: (2025)
AI-Supported Platform for System Monitoring and Decision-Making in Nuclear Waste Management with Large Language Models
di: Chang, Dongjune, et al.
Pubblicazione: (2025)
di: Chang, Dongjune, et al.
Pubblicazione: (2025)
Deception Analysis with Artificial Intelligence: An Interdisciplinary Perspective
di: Sarkadi, Stefan
Pubblicazione: (2024)
di: Sarkadi, Stefan
Pubblicazione: (2024)
Persona Alchemy: Designing, Evaluating, and Implementing Psychologically-Grounded LLM Agents for Diverse Stakeholder Representation
di: Kim, Sola, et al.
Pubblicazione: (2025)
di: Kim, Sola, et al.
Pubblicazione: (2025)
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
di: Cobben, Pepijn, et al.
Pubblicazione: (2026)
di: Cobben, Pepijn, et al.
Pubblicazione: (2026)
Investigating Tax Evasion Emergence Using Dual Large Language Model and Deep Reinforcement Learning Powered Agent-based Simulation
di: Lazebnik, Teddy, et al.
Pubblicazione: (2025)
di: Lazebnik, Teddy, et al.
Pubblicazione: (2025)
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models
di: Choi, Younwoo, et al.
Pubblicazione: (2025)
di: Choi, Younwoo, et al.
Pubblicazione: (2025)
Social Theory Should Be a Structural Prior for Agentic AI: A Formal Framework for Multi-Agent Social Systems
di: Ng, Lynnette Hui Xian, et al.
Pubblicazione: (2026)
di: Ng, Lynnette Hui Xian, et al.
Pubblicazione: (2026)
An agent-based model of modal choice with perception biases and habits
di: Adam, Carole, et al.
Pubblicazione: (2024)
di: Adam, Carole, et al.
Pubblicazione: (2024)
Insured Agents: A Decentralized Trust Insurance Mechanism for Agentic Economy
di: Hu, Botao 'Amber', et al.
Pubblicazione: (2025)
di: Hu, Botao 'Amber', et al.
Pubblicazione: (2025)
Quantifying the Lifelong Impact of Resilience Interventions via Agent-Based LLM Simulation
di: Ming, Vivienne L'Ecuyer
Pubblicazione: (2025)
di: Ming, Vivienne L'Ecuyer
Pubblicazione: (2025)
Synergy: A Next-Generation General-Purpose Agent for Open Agentic Web
di: Nie, Xiaohang, et al.
Pubblicazione: (2026)
di: Nie, Xiaohang, et al.
Pubblicazione: (2026)
A survey about perceptions of mobility to inform an agent-based simulator of subjective modal choice
di: Adam, Carole, et al.
Pubblicazione: (2025)
di: Adam, Carole, et al.
Pubblicazione: (2025)
A survey to measure cognitive biases influencing mobility choices
di: Adam, Carole
Pubblicazione: (2024)
di: Adam, Carole
Pubblicazione: (2024)
Steve: LLM Powered ChatBot for Career Progression
di: Renji, Naveen Mathews, et al.
Pubblicazione: (2025)
di: Renji, Naveen Mathews, et al.
Pubblicazione: (2025)
On the Transition to an Auction-based Intelligent Parking Assignment System
di: Alekszejenkó, Levente, et al.
Pubblicazione: (2026)
di: Alekszejenkó, Levente, et al.
Pubblicazione: (2026)
An agent-based epidemics simulation to compare and explain screening and vaccination prioritisation strategies
di: Adam, Carole, et al.
Pubblicazione: (2022)
di: Adam, Carole, et al.
Pubblicazione: (2022)
Assessing the Effects of Container Handling Strategies on Enhancing Freight Throughput
di: Rattanakunuprakarn, Sarita, et al.
Pubblicazione: (2024)
di: Rattanakunuprakarn, Sarita, et al.
Pubblicazione: (2024)
All Models Are Wrong, But Can They Be Useful? Lessons from COVID-19 Agent-Based Models: A Systematic Review
di: Von Hoene, Emma, et al.
Pubblicazione: (2025)
di: Von Hoene, Emma, et al.
Pubblicazione: (2025)
APS: Bias-Controlled Adaptive Prototype Simulation for Population-Scale LLM Agents
di: Zheng, Quan, et al.
Pubblicazione: (2026)
di: Zheng, Quan, et al.
Pubblicazione: (2026)
A Survey on Trustworthy LLM Agents: Threats and Countermeasures
di: Yu, Miao, et al.
Pubblicazione: (2025)
di: Yu, Miao, et al.
Pubblicazione: (2025)
Different Facets for Different Experts: A Framework for Streamlining The Integration of Qualitative Insights into ABM Development
di: Nallur, Vivek, et al.
Pubblicazione: (2024)
di: Nallur, Vivek, et al.
Pubblicazione: (2024)
Is Your LLM-as-a-Recommender Agent Trustable? LLMs' Recommendation is Easily Hacked by Biases (Preferences)
di: Tang, Zichen, et al.
Pubblicazione: (2026)
di: Tang, Zichen, et al.
Pubblicazione: (2026)
Votiverse: A Configurable Governance Platform for Democratic Decision-Making
di: Macrini, Diego
Pubblicazione: (2026)
di: Macrini, Diego
Pubblicazione: (2026)
Are LLM Agents the New RPA? A Comparative Study with RPA Across Enterprise Workflows
di: Průcha, Petr, et al.
Pubblicazione: (2025)
di: Průcha, Petr, et al.
Pubblicazione: (2025)
Delegations as Adaptive Representation Patterns: Rethinking Influence in Liquid Democracy
di: Grossi, Davide, et al.
Pubblicazione: (2025)
di: Grossi, Davide, et al.
Pubblicazione: (2025)
How Routing Strategies Impact Urban Emissions
di: Cornacchia, Giuliano, et al.
Pubblicazione: (2022)
di: Cornacchia, Giuliano, et al.
Pubblicazione: (2022)
Using Feasible Action-Space Reduction by Groups to fill Causal Responsibility Gaps in Spatial Interactions
di: Guenov, Vassil, et al.
Pubblicazione: (2026)
di: Guenov, Vassil, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Toward Inclusive Educational AI: Auditing Frontier LLMs through a Multiplexity Lens
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025) -
Harnessing Multi-Agent LLMs for Complex Engineering Problem-Solving: A Framework for Senior Design Projects
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025) -
Can LLMs Write Faithfully? An Agent-Based Evaluation of LLM-generated Islamic Content
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025) -
Can Agents Judge Systematic Reviews Like Humans? Evaluating SLRs with LLM-based Multi-Agent System
di: Mushtaq, Abdullah, et al.
Pubblicazione: (2025) -
IslamicLegalBench: Evaluating LLMs Knowledge and Reasoning of Islamic Law Across 1,200 Years of Islamic Pluralist Legal Traditions
di: Elmahjub, Ezieddin, et al.
Pubblicazione: (2026)