Sparks of Rationality: Do Reasoning LLMs Align with Human Judgment and Choice?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tak, Ala N., Banayeeanzade, Amin, Bolourani, Anahita, Bahrani, Fatemeh, Chaubey, Ashutosh, Karimireddy, Sai Praneeth, Schwarz, Norbert, Gratch, Jonathan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Psychological Steering in LLMs: An Evaluation of Effectiveness and Trustworthiness
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
Mechanistic Interpretability of Emotion Inference in Large Language Models
von: Tak, Ala N., et al.
Veröffentlicht: (2025)
von: Tak, Ala N., et al.
Veröffentlicht: (2025)
Sampling More, Getting Less: Calibration is the Diversity Bottleneck in LLMs
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
GPT-4 Emulates Average-Human Emotional Cognition from a Third-Person Perspective
von: Tak, Ala N., et al.
Veröffentlicht: (2024)
von: Tak, Ala N., et al.
Veröffentlicht: (2024)
GABRIL: Gaze-Based Regularization for Mitigating Causal Confusion in Imitation Learning
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
ContextLeak: Auditing Leakage in Private In-Context Learning Methods
von: Choi, Jacob, et al.
Veröffentlicht: (2025)
von: Choi, Jacob, et al.
Veröffentlicht: (2025)
Do Data Valuations Make Good Data Prices?
von: Fan, Dongyang, et al.
Veröffentlicht: (2025)
von: Fan, Dongyang, et al.
Veröffentlicht: (2025)
AutoFocus-IL: VLM-based Saliency Maps for Data-Efficient Visual Imitation Learning without Extra Human Annotations
von: Gong, Litian, et al.
Veröffentlicht: (2025)
von: Gong, Litian, et al.
Veröffentlicht: (2025)
Optimization with Access to Auxiliary Information
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2022)
von: Chayti, El Mahdi, et al.
Veröffentlicht: (2022)
A Systematic Analysis of Base Model Choice for Reward Modeling
von: Ahrabian, Kian, et al.
Veröffentlicht: (2025)
von: Ahrabian, Kian, et al.
Veröffentlicht: (2025)
EPSVec: Efficient and Private Synthetic Data Generation via Dataset Vectors
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026)
On the Limits of Momentum in Decentralized and Federated Optimization
von: Zaccone, Riccardo, et al.
Veröffentlicht: (2025)
von: Zaccone, Riccardo, et al.
Veröffentlicht: (2025)
Ghosted Layers: Unconstrained Activation Alignment for Recovering Layer-Pruned LLMs
von: Yun, Vincent-Daniel, et al.
Veröffentlicht: (2026)
von: Yun, Vincent-Daniel, et al.
Veröffentlicht: (2026)
Hybrid Learners Do Not Forget: A Brain-Inspired Neuro-Symbolic Approach to Continual Learning
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)
Do Audio LLMs Listen or Read? Analyzing and Mitigating Paralinguistic Failures with VoxParadox
von: Pang, Jiacheng, et al.
Veröffentlicht: (2026)
von: Pang, Jiacheng, et al.
Veröffentlicht: (2026)
Collaborative Heterogeneous Causal Inference Beyond Meta-analysis
von: Guo, Tianyu, et al.
Veröffentlicht: (2024)
von: Guo, Tianyu, et al.
Veröffentlicht: (2024)
Defection-Free Collaboration between Competitors in a Learning System
von: Werner, Mariel, et al.
Veröffentlicht: (2024)
von: Werner, Mariel, et al.
Veröffentlicht: (2024)
Robust Multi-Agent LLMs under Byzantine Faults
von: Lee, Haejoon, et al.
Veröffentlicht: (2026)
von: Lee, Haejoon, et al.
Veröffentlicht: (2026)
LIA: Privacy-Preserving Data Quality Evaluation in Federated Learning Using a Lazy Influence Approximation
von: Rokvic, Ljubomir, et al.
Veröffentlicht: (2022)
von: Rokvic, Ljubomir, et al.
Veröffentlicht: (2022)
Communication-Efficient Heterogeneous Federated Learning with Generalized Heavy-Ball Momentum
von: Zaccone, Riccardo, et al.
Veröffentlicht: (2023)
von: Zaccone, Riccardo, et al.
Veröffentlicht: (2023)
Beyond URLs: Metadata Diversity and Position for Efficient LLM Pretraining
von: Fan, Dongyang, et al.
Veröffentlicht: (2025)
von: Fan, Dongyang, et al.
Veröffentlicht: (2025)
A Differentially Private Kaplan-Meier Estimator for Privacy-Preserving Survival Analysis
von: Veeraragavan, Narasimha Raghavan, et al.
Veröffentlicht: (2024)
von: Veeraragavan, Narasimha Raghavan, et al.
Veröffentlicht: (2024)
DAVED: Data Acquisition via Experimental Design for Data Markets
von: Lu, Charles, et al.
Veröffentlicht: (2024)
von: Lu, Charles, et al.
Veröffentlicht: (2024)
Hair-Trigger Alignment: Black-Box Evaluation Cannot Guarantee Post-Update Alignment
von: Bakman, Yavuz, et al.
Veröffentlicht: (2026)
von: Bakman, Yavuz, et al.
Veröffentlicht: (2026)
MoD-DPO: Towards Mitigating Cross-modal Hallucinations in Omni LLMs using Modality Decoupled Preference Optimization
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2026)
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2026)
Entropy-driven Fair and Effective Federated Learning
von: Wang, Lin, et al.
Veröffentlicht: (2023)
von: Wang, Lin, et al.
Veröffentlicht: (2023)
VoxGuard: Evaluating User and Attribute Privacy in Speech via Membership Inference Attacks
von: Tsaprazlis, Efthymios, et al.
Veröffentlicht: (2025)
von: Tsaprazlis, Efthymios, et al.
Veröffentlicht: (2025)
Privacy Can Arise Endogenously in an Economic System with Learning Agents
von: Ananthakrishnan, Nivasini, et al.
Veröffentlicht: (2024)
von: Ananthakrishnan, Nivasini, et al.
Veröffentlicht: (2024)
Conformal Prediction Adaptive to Unknown Subpopulation Shifts
von: Wang, Nien-Shao, et al.
Veröffentlicht: (2025)
von: Wang, Nien-Shao, et al.
Veröffentlicht: (2025)
f-INE: A Hypothesis Testing Framework for Estimating Influence under Training Randomness
von: Panda, Subhodip, et al.
Veröffentlicht: (2025)
von: Panda, Subhodip, et al.
Veröffentlicht: (2025)
AVERE: Improving Audiovisual Emotion Reasoning with Preference Optimization
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2026)
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2026)
Rethinking Visual Privacy: A Compositional Privacy Risk Framework for Severity Assessment with VLMs
von: Tsaprazlis, Efthymios, et al.
Veröffentlicht: (2026)
von: Tsaprazlis, Efthymios, et al.
Veröffentlicht: (2026)
A Closer Look at Personalized Fine-Tuning in Heterogeneous Federated Learning
von: Chen, Minghui, et al.
Veröffentlicht: (2025)
von: Chen, Minghui, et al.
Veröffentlicht: (2025)
Theoretical Insights into Overparameterized Models in Multi-Task and Replay-Based Continual Learning
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2024)
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2024)
Reasons to Reject? Aligning Language Models with Judgments
von: Xu, Weiwen, et al.
Veröffentlicht: (2023)
von: Xu, Weiwen, et al.
Veröffentlicht: (2023)
Face-LLaVA: Facial Expression and Attribute Understanding through Instruction Tuning
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2025)
von: Chaubey, Ashutosh, et al.
Veröffentlicht: (2025)
Salience Adjustment for Context-Based Emotion Recognition
von: Han, Bin, et al.
Veröffentlicht: (2025)
von: Han, Bin, et al.
Veröffentlicht: (2025)
OpaqueToolsBench: Learning Nuances of Tool Behavior Through Interaction
von: Hallinan, Skyler, et al.
Veröffentlicht: (2026)
von: Hallinan, Skyler, et al.
Veröffentlicht: (2026)
Sparks of Cooperative Reasoning: LLMs as Strategic Hanabi Agents
von: Ramesh, Mahesh, et al.
Veröffentlicht: (2026)
von: Ramesh, Mahesh, et al.
Veröffentlicht: (2026)
DialogueReason: Rule-Based RL Sparks Dialogue Reasoning in LLMs
von: Shu, Yubo, et al.
Veröffentlicht: (2025)
von: Shu, Yubo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Psychological Steering in LLMs: An Evaluation of Effectiveness and Trustworthiness
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025) -
Mechanistic Interpretability of Emotion Inference in Large Language Models
von: Tak, Ala N., et al.
Veröffentlicht: (2025) -
Sampling More, Getting Less: Calibration is the Diversity Bottleneck in LLMs
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2026) -
GPT-4 Emulates Average-Human Emotional Cognition from a Third-Person Perspective
von: Tak, Ala N., et al.
Veröffentlicht: (2024) -
GABRIL: Gaze-Based Regularization for Mitigating Causal Confusion in Imitation Learning
von: Banayeeanzade, Amin, et al.
Veröffentlicht: (2025)