PERSPECTRA: A Scalable and Configurable Pluralist Benchmark of Perspectives from Arguments
Fuente:
arXiv
Guardado en:
| Autores principales: | Nie, Shangrui, Omoomi, Kian, Flek, Lucie, Zhao, Zhixue, Welch, Charles |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Survey-to-Behavior: Downstream Alignment of Human Values in LLMs via Survey Questions
por: Nie, Shangrui, et al.
Publicado: (2025)
por: Nie, Shangrui, et al.
Publicado: (2025)
A Critical Reflection and Forward Perspective on Empathy and Natural Language Processing
por: Lahnala, Allison, et al.
Publicado: (2022)
por: Lahnala, Allison, et al.
Publicado: (2022)
Examining the Utility of Self-disclosure Types for Modeling Annotators of Social Norms
por: Henderson, Kieran, et al.
Publicado: (2025)
por: Henderson, Kieran, et al.
Publicado: (2025)
Funzac at CoMeDi Shared Task: Modeling Annotator Disagreement from Word-In-Context Perspectives
por: Sarumi, Olufunke O., et al.
Publicado: (2025)
por: Sarumi, Olufunke O., et al.
Publicado: (2025)
The Muddy Waters of Modeling Empathy in Language: The Practical Impacts of Theoretical Constructs
por: Lahnala, Allison, et al.
Publicado: (2025)
por: Lahnala, Allison, et al.
Publicado: (2025)
Do Multilingual Large Language Models Mitigate Stereotype Bias?
por: Nie, Shangrui, et al.
Publicado: (2024)
por: Nie, Shangrui, et al.
Publicado: (2024)
On the Limitations of Language Targeted Pruning: Investigating the Calibration Language Impact in Multilingual LLM Pruning
por: Kurz, Simon, et al.
Publicado: (2024)
por: Kurz, Simon, et al.
Publicado: (2024)
Do LLMs Provide Consistent Answers to Health-Related Questions across Languages?
por: Schlicht, Ipek Baris, et al.
Publicado: (2025)
por: Schlicht, Ipek Baris, et al.
Publicado: (2025)
ISCA: A Framework for Interview-Style Conversational Agents
por: Welch, Charles, et al.
Publicado: (2025)
por: Welch, Charles, et al.
Publicado: (2025)
Disparities in Multilingual LLM-Based Healthcare Q&A
por: Schlicht, Ipek Baris, et al.
Publicado: (2025)
por: Schlicht, Ipek Baris, et al.
Publicado: (2025)
Encoder Fine-tuning with Stochastic Sampling Outperforms Open-weight GPT in Astronomy Knowledge Extraction
por: Rawat, Shivam, et al.
Publicado: (2025)
por: Rawat, Shivam, et al.
Publicado: (2025)
Exploring Robustness of LLMs to Paraphrasing Based on Sociodemographic Factors
por: Arora, Pulkit, et al.
Publicado: (2025)
por: Arora, Pulkit, et al.
Publicado: (2025)
Exploring Robustness of Multilingual LLMs on Real-World Noisy Data
por: Aliakbarzadeh, Amirhossein, et al.
Publicado: (2025)
por: Aliakbarzadeh, Amirhossein, et al.
Publicado: (2025)
Corpus Considerations for Annotator Modeling and Scaling
por: Sarumi, Olufunke O., et al.
Publicado: (2024)
por: Sarumi, Olufunke O., et al.
Publicado: (2024)
IKnow: Instruction-Knowledge-Aware Continual Pretraining for Effective Domain Adaptation
por: Zhang, Tianyi, et al.
Publicado: (2025)
por: Zhang, Tianyi, et al.
Publicado: (2025)
Multi-Hop Reasoning for Question Answering with Hyperbolic Representations
por: Welz, Simon, et al.
Publicado: (2025)
por: Welz, Simon, et al.
Publicado: (2025)
Probing the Robustness of Theory of Mind in Large Language Models
por: Nickel, Christian, et al.
Publicado: (2024)
por: Nickel, Christian, et al.
Publicado: (2024)
Label-Consistent Data Generation for Aspect-Based Sentiment Analysis Using LLM Agents
por: Monfared, Mohammad H. A., et al.
Publicado: (2026)
por: Monfared, Mohammad H. A., et al.
Publicado: (2026)
ArithmAttack: Evaluating Robustness of LLMs to Noisy Context in Math Problem Solving
por: Abedin, Zain Ul, et al.
Publicado: (2025)
por: Abedin, Zain Ul, et al.
Publicado: (2025)
Understanding Artificial Theory of Mind: Perturbed Tasks and Reasoning in Large Language Models
por: Nickel, Christian, et al.
Publicado: (2026)
por: Nickel, Christian, et al.
Publicado: (2026)
More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persists
por: Alavi, Khashayar, et al.
Publicado: (2025)
por: Alavi, Khashayar, et al.
Publicado: (2025)
Reasoning Primitives in Hybrid and Non-Hybrid LLMs: Do Architectural Differences Yield Advantages in State-Tracking and Recall?
por: Rawat, Shivam, et al.
Publicado: (2026)
por: Rawat, Shivam, et al.
Publicado: (2026)
Pitfalls of Conversational LLMs on News Debiasing
por: Schlicht, Ipek Baris, et al.
Publicado: (2024)
por: Schlicht, Ipek Baris, et al.
Publicado: (2024)
Reinforcement Learning Amplifies Emergent Misalignment from Harmless Rewards
por: Jørgenvåg, Magnus, et al.
Publicado: (2026)
por: Jørgenvåg, Magnus, et al.
Publicado: (2026)
Can LLM Agents Identify Spoken Dialects like a Linguist?
por: Bystrich, Tobias, et al.
Publicado: (2026)
por: Bystrich, Tobias, et al.
Publicado: (2026)
PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media
por: Kachwala, Zoher, et al.
Publicado: (2026)
por: Kachwala, Zoher, et al.
Publicado: (2026)
Fine-Grained Perspectives: Modeling Explanations with Annotator-Specific Rationales
por: Sarumi, Olufunke O., et al.
Publicado: (2026)
por: Sarumi, Olufunke O., et al.
Publicado: (2026)
VITAL: A New Dataset for Benchmarking Pluralistic Alignment in Healthcare
por: Shetty, Anudeex, et al.
Publicado: (2025)
por: Shetty, Anudeex, et al.
Publicado: (2025)
Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models
por: Zhao, Zhixue, et al.
Publicado: (2024)
por: Zhao, Zhixue, et al.
Publicado: (2024)
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models
por: Zhao, Zhixue, et al.
Publicado: (2024)
por: Zhao, Zhixue, et al.
Publicado: (2024)
PERSONA: A Reproducible Testbed for Pluralistic Alignment
por: Castricato, Louis, et al.
Publicado: (2024)
por: Castricato, Louis, et al.
Publicado: (2024)
Unifying the Extremes: Developing a Unified Model for Detecting and Predicting Extremist Traits and Radicalization
por: Lahnala, Allison, et al.
Publicado: (2025)
por: Lahnala, Allison, et al.
Publicado: (2025)
Can Stories Help LLMs Reason? Curating Information Space Through Narrative
por: Javadi, Vahid Sadiri, et al.
Publicado: (2024)
por: Javadi, Vahid Sadiri, et al.
Publicado: (2024)
USDC: A Dataset of $\underline{U}$ser $\underline{S}$tance and $\underline{D}$ogmatism in Long $\underline{C}$onversations
por: Marreddy, Mounika, et al.
Publicado: (2024)
por: Marreddy, Mounika, et al.
Publicado: (2024)
Improving Low-Resource Dialect Classification Using Retrieval-based Voice Conversion
por: Fischbach, Lea, et al.
Publicado: (2025)
por: Fischbach, Lea, et al.
Publicado: (2025)
Explanation Generation for Contradiction Reconciliation with LLMs
por: Chan, Jason, et al.
Publicado: (2026)
por: Chan, Jason, et al.
Publicado: (2026)
Position: Logical Soundness is not a Reliable Criterion for Neurosymbolic Fact-Checking with LLMs
por: Chan, Jason, et al.
Publicado: (2026)
por: Chan, Jason, et al.
Publicado: (2026)
Label Set Optimization via Activation Distribution Kurtosis for Zero-shot Classification with Generative Models
por: Li, Yue, et al.
Publicado: (2024)
por: Li, Yue, et al.
Publicado: (2024)
RULEBREAKERS: Challenging LLMs at the Crossroads between Formal Logic and Human-like Reasoning
por: Chan, Jason, et al.
Publicado: (2024)
por: Chan, Jason, et al.
Publicado: (2024)
Position: On the Methodological Pitfalls of Evaluating Base LLMs for Reasoning
por: Chan, Jason, et al.
Publicado: (2025)
por: Chan, Jason, et al.
Publicado: (2025)
Ejemplares similares
-
Survey-to-Behavior: Downstream Alignment of Human Values in LLMs via Survey Questions
por: Nie, Shangrui, et al.
Publicado: (2025) -
A Critical Reflection and Forward Perspective on Empathy and Natural Language Processing
por: Lahnala, Allison, et al.
Publicado: (2022) -
Examining the Utility of Self-disclosure Types for Modeling Annotators of Social Norms
por: Henderson, Kieran, et al.
Publicado: (2025) -
Funzac at CoMeDi Shared Task: Modeling Annotator Disagreement from Word-In-Context Perspectives
por: Sarumi, Olufunke O., et al.
Publicado: (2025) -
The Muddy Waters of Modeling Empathy in Language: The Practical Impacts of Theoretical Constructs
por: Lahnala, Allison, et al.
Publicado: (2025)