LocalValueBench: A Collaboratively Built and Extensible Benchmark for Evaluating Localized Value Alignment and Ethical Safety in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Meadows, Gwenyth Isobel, Lau, Nicholas Wai Long, Susanto, Eva Adelina, Yu, Chi Lok, Paul, Aditya |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AgentPeerTalk: Empowering Students through Agentic-AI-Driven Discernment of Bullying and Joking in Peer Interactions in Schools
von: Paul, Aditya, et al.
Veröffentlicht: (2024)
von: Paul, Aditya, et al.
Veröffentlicht: (2024)
Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values?
von: Shen, Hua, et al.
Veröffentlicht: (2025)
von: Shen, Hua, et al.
Veröffentlicht: (2025)
EigenBench: A Comparative Behavioral Measure of Value Alignment
von: Chang, Jonathn, et al.
Veröffentlicht: (2025)
von: Chang, Jonathn, et al.
Veröffentlicht: (2025)
WorldValuesBench: A Large-Scale Benchmark Dataset for Multi-Cultural Value Awareness of Language Models
von: Zhao, Wenlong, et al.
Veröffentlicht: (2024)
von: Zhao, Wenlong, et al.
Veröffentlicht: (2024)
AdAEM: An Adaptively and Automated Extensible Measurement of LLMs' Value Difference
von: Yao, Jing, et al.
Veröffentlicht: (2025)
von: Yao, Jing, et al.
Veröffentlicht: (2025)
LocalBench: Benchmarking LLMs on County-Level Local Knowledge and Reasoning
von: Gao, Zihan, et al.
Veröffentlicht: (2025)
von: Gao, Zihan, et al.
Veröffentlicht: (2025)
Flames: Benchmarking Value Alignment of LLMs in Chinese
von: Huang, Kexin, et al.
Veröffentlicht: (2023)
von: Huang, Kexin, et al.
Veröffentlicht: (2023)
Diverse Human Value Alignment for Large Language Models via Ethical Reasoning
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
von: Wang, Jiahao, et al.
Veröffentlicht: (2025)
LLMs Homogenize Values in Constructive Arguments on Value-Laden Topics
von: Shahid, Farhana, et al.
Veröffentlicht: (2025)
von: Shahid, Farhana, et al.
Veröffentlicht: (2025)
An Evaluation of Cultural Value Alignment in LLM
von: Sukiennik, Nicholas, et al.
Veröffentlicht: (2025)
von: Sukiennik, Nicholas, et al.
Veröffentlicht: (2025)
ValueCompass: A Framework for Measuring Contextual Value Alignment Between Human and LLMs
von: Shen, Hua, et al.
Veröffentlicht: (2024)
von: Shen, Hua, et al.
Veröffentlicht: (2024)
Dynamic Normativity: Necessary and Sufficient Conditions for Value Alignment
von: Corrêa, Nicholas Kluge
Veröffentlicht: (2024)
von: Corrêa, Nicholas Kluge
Veröffentlicht: (2024)
Ethical Reasoning and Moral Value Alignment of LLMs Depend on the Language we Prompt them in
von: Agarwal, Utkarsh, et al.
Veröffentlicht: (2024)
von: Agarwal, Utkarsh, et al.
Veröffentlicht: (2024)
Deep Learning-Based BMD Estimation from Radiographs with Conformal Uncertainty Quantification
von: Hui, Long, et al.
Veröffentlicht: (2025)
von: Hui, Long, et al.
Veröffentlicht: (2025)
Evaluating AI Alignment in LLMs: Output Analysis of Value Priorities Across 75 Models with Human Benchmarking
von: Lau, Gabriel Rongyang, et al.
Veröffentlicht: (2025)
von: Lau, Gabriel Rongyang, et al.
Veröffentlicht: (2025)
Local-Prompt: Extensible Local Prompts for Few-Shot Out-of-Distribution Detection
von: Zeng, Fanhu, et al.
Veröffentlicht: (2024)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2024)
UAL-Bench: The First Comprehensive Unusual Activity Localization Benchmark
von: Abdullah, Hasnat Md, et al.
Veröffentlicht: (2024)
von: Abdullah, Hasnat Md, et al.
Veröffentlicht: (2024)
VAL-Bench: Belief Consistency as a measure for Value Alignment in Language Models
von: Gupta, Aman, et al.
Veröffentlicht: (2025)
von: Gupta, Aman, et al.
Veröffentlicht: (2025)
Beyond Single-Sentence Prompts: Upgrading Value Alignment Benchmarks with Dialogues and Stories
von: Zhang, Yazhou, et al.
Veröffentlicht: (2025)
von: Zhang, Yazhou, et al.
Veröffentlicht: (2025)
Value Drifts: Tracing Value Alignment During LLM Post-Training
von: Bhatia, Mehar, et al.
Veröffentlicht: (2025)
von: Bhatia, Mehar, et al.
Veröffentlicht: (2025)
Benchmarking Multi-National Value Alignment for Large Language Models
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
von: Shi, Weijie, et al.
Veröffentlicht: (2025)
Value Alignment Tax: Measuring Value Trade-offs in LLM Alignment
von: Chen, Jiajun, et al.
Veröffentlicht: (2026)
von: Chen, Jiajun, et al.
Veröffentlicht: (2026)
KorNAT: LLM Alignment Benchmark for Korean Social Values and Common Knowledge
von: Lee, Jiyoung, et al.
Veröffentlicht: (2024)
von: Lee, Jiyoung, et al.
Veröffentlicht: (2024)
ValueBench: Towards Comprehensively Evaluating Value Orientations and Understanding of Large Language Models
von: Ren, Yuanyi, et al.
Veröffentlicht: (2024)
von: Ren, Yuanyi, et al.
Veröffentlicht: (2024)
Can LLMs Grasp Implicit Cultural Values? Benchmarking LLMs' Cultural Intelligence with CQ-Bench
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
Exploring the Impact of AI Value Alignment in Collaborative Ideation: Effects on Perception, Ownership, and Output
von: Guo, Alicia, et al.
Veröffentlicht: (2024)
von: Guo, Alicia, et al.
Veröffentlicht: (2024)
PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay
von: Khetan, Rohan, et al.
Veröffentlicht: (2026)
von: Khetan, Rohan, et al.
Veröffentlicht: (2026)
Video-SafetyBench: A Benchmark for Safety Evaluation of Video LVLMs
von: Liu, Xuannan, et al.
Veröffentlicht: (2025)
von: Liu, Xuannan, et al.
Veröffentlicht: (2025)
Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2026)
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2026)
SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment
von: Li, Hao, et al.
Veröffentlicht: (2026)
von: Li, Hao, et al.
Veröffentlicht: (2026)
DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions
von: Wang, Xinran, et al.
Veröffentlicht: (2026)
von: Wang, Xinran, et al.
Veröffentlicht: (2026)
An Extensible Framework for Open Heterogeneous Collaborative Perception
von: Lu, Yifan, et al.
Veröffentlicht: (2024)
von: Lu, Yifan, et al.
Veröffentlicht: (2024)
Detect2Interact: Localizing Object Key Field in Visual Question Answering (VQA) with LLMs
von: Wang, Jialou, et al.
Veröffentlicht: (2024)
von: Wang, Jialou, et al.
Veröffentlicht: (2024)
Deep Value Benchmark: Measuring Whether Models Generalize Deep Values or Shallow Preferences
von: Ashkinaze, Joshua, et al.
Veröffentlicht: (2025)
von: Ashkinaze, Joshua, et al.
Veröffentlicht: (2025)
LocateEdit-Bench: A Benchmark for Instruction-Based Editing Localization
von: Wu, Shiyu, et al.
Veröffentlicht: (2026)
von: Wu, Shiyu, et al.
Veröffentlicht: (2026)
The Value of Disagreement in AI Design, Evaluation, and Alignment
von: Fazelpour, Sina, et al.
Veröffentlicht: (2025)
von: Fazelpour, Sina, et al.
Veröffentlicht: (2025)
New Bayesian method for estimation of Value at Risk and Conditional Value at Risk
von: Martín, Jacinto, et al.
Veröffentlicht: (2023)
von: Martín, Jacinto, et al.
Veröffentlicht: (2023)
Applying Value Sensitive Design to Location-Based Services: Designing for Shared Spaces and Local Conditions
von: Kegalle, Hiruni, et al.
Veröffentlicht: (2026)
von: Kegalle, Hiruni, et al.
Veröffentlicht: (2026)
Value Alignment from Unstructured Text
von: Padhi, Inkit, et al.
Veröffentlicht: (2024)
von: Padhi, Inkit, et al.
Veröffentlicht: (2024)
QualBench: Benchmarking Chinese LLMs with Localized Professional Qualifications for Vertical Domain Evaluation
von: Hong, Mengze, et al.
Veröffentlicht: (2025)
von: Hong, Mengze, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AgentPeerTalk: Empowering Students through Agentic-AI-Driven Discernment of Bullying and Joking in Peer Interactions in Schools
von: Paul, Aditya, et al.
Veröffentlicht: (2024) -
Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values?
von: Shen, Hua, et al.
Veröffentlicht: (2025) -
EigenBench: A Comparative Behavioral Measure of Value Alignment
von: Chang, Jonathn, et al.
Veröffentlicht: (2025) -
WorldValuesBench: A Large-Scale Benchmark Dataset for Multi-Cultural Value Awareness of Language Models
von: Zhao, Wenlong, et al.
Veröffentlicht: (2024) -
AdAEM: An Adaptively and Automated Extensible Measurement of LLMs' Value Difference
von: Yao, Jing, et al.
Veröffentlicht: (2025)