Salvato in:
| Autori principali: | Poole-Dayan, Elinor, Wu, Jiayi, Sorensen, Taylor, Pei, Jiaxin, Bakker, Michiel A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2512.01351 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LLM Targeted Underperformance Disproportionately Impacts Vulnerable Users
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2024)
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2024)
On the Relationship between Truth and Political Bias in Language Models
di: Fulay, Suyash, et al.
Pubblicazione: (2024)
di: Fulay, Suyash, et al.
Pubblicazione: (2024)
From Delegates to Trustees: How Optimizing for Long-Term Interests Shapes Bias and Alignment in LLM
di: Fulay, Suyash, et al.
Pubblicazione: (2025)
di: Fulay, Suyash, et al.
Pubblicazione: (2025)
Do LLMs Understand Social Knowledge? Evaluating the Sociability of Large Language Models with SocKET Benchmark
di: Choi, Minje, et al.
Pubblicazione: (2023)
di: Choi, Minje, et al.
Pubblicazione: (2023)
An AI-Powered Framework for Analyzing Collective Idea Evolution in Deliberative Assemblies
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)
Value Profiles for Encoding Human Variation
di: Sorensen, Taylor, et al.
Pubblicazione: (2025)
di: Sorensen, Taylor, et al.
Pubblicazione: (2025)
Opt-ICL at LeWiDi-2025: Maximizing In-Context Signal from Rater Examples via Meta-Learning
di: Sorensen, Taylor, et al.
Pubblicazione: (2025)
di: Sorensen, Taylor, et al.
Pubblicazione: (2025)
AI Assistance Reduces Persistence and Hurts Independent Performance
di: Liu, Grace, et al.
Pubblicazione: (2026)
di: Liu, Grace, et al.
Pubblicazione: (2026)
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
di: Vaccaro, Michelle, et al.
Pubblicazione: (2026)
di: Vaccaro, Michelle, et al.
Pubblicazione: (2026)
When is using AI the rational choice? The importance of counterfactuals in AI deployment decisions
di: Lehner, Paul, et al.
Pubblicazione: (2025)
di: Lehner, Paul, et al.
Pubblicazione: (2025)
Conformal Risk Control for Safety-Critical Wildfire Evacuation Mapping: A Comparative Study of Tabular, Spatial, and Graph-Based Models
di: Dayan, Baljinnyam
Pubblicazione: (2026)
di: Dayan, Baljinnyam
Pubblicazione: (2026)
Belief Engine: Configurable and Inspectable Stance Dynamics in Multi-Agent LLM Deliberation
di: Yang, Joshua C., et al.
Pubblicazione: (2026)
di: Yang, Joshua C., et al.
Pubblicazione: (2026)
Tell Me Why: Incentivizing Explanations
di: Srinivasan, Siddarth, et al.
Pubblicazione: (2025)
di: Srinivasan, Siddarth, et al.
Pubblicazione: (2025)
Tapilot-Crossing: Benchmarking and Evolving LLMs Towards Interactive Data Analysis Agents
di: Li, Jinyang, et al.
Pubblicazione: (2024)
di: Li, Jinyang, et al.
Pubblicazione: (2024)
Overtone: Cyclic Patch Modulation for Clean, Efficient, and Flexible Physics Emulators
di: Mukhopadhyay, Payel, et al.
Pubblicazione: (2025)
di: Mukhopadhyay, Payel, et al.
Pubblicazione: (2025)
Measuring What Matters: The AI Pluralism Index
di: Mushkani, Rashid
Pubblicazione: (2025)
di: Mushkani, Rashid
Pubblicazione: (2025)
The World According to LLMs: How Geographic Origin Influences LLMs' Entity Deduction Capabilities
di: Lalai, Harsh Nishant, et al.
Pubblicazione: (2025)
di: Lalai, Harsh Nishant, et al.
Pubblicazione: (2025)
Why Isn't Relational Learning Taking Over the World?
di: Poole, David
Pubblicazione: (2025)
di: Poole, David
Pubblicazione: (2025)
RE-PO: Robust Enhanced Policy Optimization as a General Framework for LLM Alignment
di: Cao, Xiaoyang, et al.
Pubblicazione: (2025)
di: Cao, Xiaoyang, et al.
Pubblicazione: (2025)
Benchmarking for Domain-Specific LLMs: A Case Study on Academia and Beyond
di: Chen, Rubing, et al.
Pubblicazione: (2025)
di: Chen, Rubing, et al.
Pubblicazione: (2025)
DHP Benchmark: Are LLMs Good NLG Evaluators?
di: Wang, Yicheng, et al.
Pubblicazione: (2024)
di: Wang, Yicheng, et al.
Pubblicazione: (2024)
Ontology Learning with LLMs: A Benchmark Study on Axiom Identification
di: Bakker, Roos M., et al.
Pubblicazione: (2025)
di: Bakker, Roos M., et al.
Pubblicazione: (2025)
GeoEval: Benchmark for Evaluating LLMs and Multi-Modal Models on Geometry Problem-Solving
di: Zhang, Jiaxin, et al.
Pubblicazione: (2024)
di: Zhang, Jiaxin, et al.
Pubblicazione: (2024)
What Does the AI Doctor Value? Auditing Pluralism in the Clinical Ethics of Language Models
di: Chandak, Payal, et al.
Pubblicazione: (2026)
di: Chandak, Payal, et al.
Pubblicazione: (2026)
AI and Collective Decisions: Strengthening Legitimacy and Losers' Consent
di: Fulay, Suyash, et al.
Pubblicazione: (2026)
di: Fulay, Suyash, et al.
Pubblicazione: (2026)
Plurals: A System for Guiding LLMs Via Simulated Social Ensembles
di: Ashkinaze, Joshua, et al.
Pubblicazione: (2024)
di: Ashkinaze, Joshua, et al.
Pubblicazione: (2024)
Sociodemographic Prompting is Not Yet an Effective Approach for Simulating Subjective Judgments with LLMs
di: Sun, Huaman, et al.
Pubblicazione: (2023)
di: Sun, Huaman, et al.
Pubblicazione: (2023)
Beyond Face Swapping: A Diffusion-Based Digital Human Benchmark for Multimodal Deepfake Detection
di: Liu, Jiaxin, et al.
Pubblicazione: (2025)
di: Liu, Jiaxin, et al.
Pubblicazione: (2025)
Stop Automating Peer Review Without Rigorous Evaluation
di: Baumann, Joachim, et al.
Pubblicazione: (2026)
di: Baumann, Joachim, et al.
Pubblicazione: (2026)
MULTITEXTEDIT: Benchmarking Cross-Lingual Degradation in Text-in-Image Editing
di: Cheng, Liwei, et al.
Pubblicazione: (2026)
di: Cheng, Liwei, et al.
Pubblicazione: (2026)
Error-related Potential Variability: Exploring the Effects on Classification and Transferability
di: Poole, Benjamin, et al.
Pubblicazione: (2023)
di: Poole, Benjamin, et al.
Pubblicazione: (2023)
Comparing Traditional and Reinforcement-Learning Methods for Energy Storage Control
di: Ginzburg, Elinor, et al.
Pubblicazione: (2025)
di: Ginzburg, Elinor, et al.
Pubblicazione: (2025)
Beyond Binary Moral Judgment: Modeling Ethical Pluralism in AI
di: Aijaz, Aisha, et al.
Pubblicazione: (2026)
di: Aijaz, Aisha, et al.
Pubblicazione: (2026)
MolGround: A Benchmark for Molecular Grounding
di: Wu, Jiaxin, et al.
Pubblicazione: (2025)
di: Wu, Jiaxin, et al.
Pubblicazione: (2025)
KNVQA: A Benchmark for evaluation knowledge-based VQA
di: Cheng, Sirui, et al.
Pubblicazione: (2023)
di: Cheng, Sirui, et al.
Pubblicazione: (2023)
Benchmark Health Index: A Systematic Framework for Benchmarking the Benchmarks of LLMs
di: Zhu, Longyuan, et al.
Pubblicazione: (2026)
di: Zhu, Longyuan, et al.
Pubblicazione: (2026)
HugAgent: Benchmarking LLMs for Simulation of Individualized Human Reasoning
di: Li, Chance Jiajie, et al.
Pubblicazione: (2025)
di: Li, Chance Jiajie, et al.
Pubblicazione: (2025)
FIMA-Q: Post-Training Quantization for Vision Transformers by Fisher Information Matrix Approximation
di: Wu, Zhuguanyu, et al.
Pubblicazione: (2025)
di: Wu, Zhuguanyu, et al.
Pubblicazione: (2025)
Steerable Pluralism: Pluralistic Alignment via Few-Shot Comparative Regression
di: Adams, Jadie, et al.
Pubblicazione: (2025)
di: Adams, Jadie, et al.
Pubblicazione: (2025)
Towards Cross-lingual Values Judgment: A Consensus-Pluralism Perspective
di: Chen, Yukun, et al.
Pubblicazione: (2026)
di: Chen, Yukun, et al.
Pubblicazione: (2026)
Documenti analoghi
-
LLM Targeted Underperformance Disproportionately Impacts Vulnerable Users
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2024) -
On the Relationship between Truth and Political Bias in Language Models
di: Fulay, Suyash, et al.
Pubblicazione: (2024) -
From Delegates to Trustees: How Optimizing for Long-Term Interests Shapes Bias and Alignment in LLM
di: Fulay, Suyash, et al.
Pubblicazione: (2025) -
Do LLMs Understand Social Knowledge? Evaluating the Sociability of Large Language Models with SocKET Benchmark
di: Choi, Minje, et al.
Pubblicazione: (2023) -
An AI-Powered Framework for Analyzing Collective Idea Evolution in Deliberative Assemblies
di: Poole-Dayan, Elinor, et al.
Pubblicazione: (2025)