Relative Value Biases in Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hayes, William M., Yax, Nicolas, Palminteri, Stefano |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Large Language Models are Biased Reinforcement Learners
von: Hayes, William M., et al.
Veröffentlicht: (2024)
von: Hayes, William M., et al.
Veröffentlicht: (2024)
LogProber: Disentangling confidence from contamination in LLM responses
von: Yax, Nicolas, et al.
Veröffentlicht: (2024)
von: Yax, Nicolas, et al.
Veröffentlicht: (2024)
PhyloLM : Inferring the Phylogeny of Large Language Models and Predicting their Performances in Benchmarks
von: Yax, Nicolas, et al.
Veröffentlicht: (2024)
von: Yax, Nicolas, et al.
Veröffentlicht: (2024)
Large Language Models are Geographically Biased
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
von: Manvi, Rohin, et al.
Veröffentlicht: (2024)
Active Preference Learning for Large Language Models
von: Muldrew, William, et al.
Veröffentlicht: (2024)
von: Muldrew, William, et al.
Veröffentlicht: (2024)
Do Large Language Models Show Biases in Causal Learning?
von: Carro, Maria Victoria, et al.
Veröffentlicht: (2024)
von: Carro, Maria Victoria, et al.
Veröffentlicht: (2024)
BiasJailbreak:Analyzing Ethical Biases and Jailbreak Vulnerabilities in Large Language Models
von: Lee, Isack, et al.
Veröffentlicht: (2024)
von: Lee, Isack, et al.
Veröffentlicht: (2024)
Reward Models Inherit Value Biases from Pretraining
von: Christian, Brian, et al.
Veröffentlicht: (2026)
von: Christian, Brian, et al.
Veröffentlicht: (2026)
FairPy: A Toolkit for Evaluation of Prediction Biases and their Mitigation in Large Language Models
von: Viswanath, Hrishikesh, et al.
Veröffentlicht: (2023)
von: Viswanath, Hrishikesh, et al.
Veröffentlicht: (2023)
Can Small-Scale Data Poisoning Exacerbate Dialect-Linked Biases in Large Language Models?
von: Abbas, Chaymaa, et al.
Veröffentlicht: (2025)
von: Abbas, Chaymaa, et al.
Veröffentlicht: (2025)
Internal Value Alignment in Large Language Models through Controlled Value Vector Activation
von: Jin, Haoran, et al.
Veröffentlicht: (2025)
von: Jin, Haoran, et al.
Veröffentlicht: (2025)
Large Language Models are Good Relational Learners
von: Wu, Fang, et al.
Veröffentlicht: (2025)
von: Wu, Fang, et al.
Veröffentlicht: (2025)
Uncovering Biases with Reflective Large Language Models
von: Chang, Edward Y.
Veröffentlicht: (2024)
von: Chang, Edward Y.
Veröffentlicht: (2024)
Explaining Large Language Models Decisions Using Shapley Values
von: Mohammadi, Behnam
Veröffentlicht: (2024)
von: Mohammadi, Behnam
Veröffentlicht: (2024)
Introducing Background Temperature to Characterise Hidden Randomness in Large Language Models
von: Messina, Alberto, et al.
Veröffentlicht: (2026)
von: Messina, Alberto, et al.
Veröffentlicht: (2026)
PoliTune: Analyzing the Impact of Data Selection and Fine-Tuning on Economic and Political Biases in Large Language Models
von: Agiza, Ahmed, et al.
Veröffentlicht: (2024)
von: Agiza, Ahmed, et al.
Veröffentlicht: (2024)
Reframing Data Value for Large Language Models Through the Lens of Plausibility
von: Rammal, Mohamad Rida, et al.
Veröffentlicht: (2024)
von: Rammal, Mohamad Rida, et al.
Veröffentlicht: (2024)
Regurgitative Training: The Value of Real Data in Training Large Language Models
von: Zhang, Jinghui, et al.
Veröffentlicht: (2024)
von: Zhang, Jinghui, et al.
Veröffentlicht: (2024)
Between Circuits and Chomsky: Pre-pretraining on Formal Languages Imparts Linguistic Biases
von: Hu, Michael Y., et al.
Veröffentlicht: (2025)
von: Hu, Michael Y., et al.
Veröffentlicht: (2025)
Benchmarking Cognitive Biases in Large Language Models as Evaluators
von: Koo, Ryan, et al.
Veröffentlicht: (2023)
von: Koo, Ryan, et al.
Veröffentlicht: (2023)
Empowering Many, Biasing a Few: Generalist Credit Scoring through Large Language Models
von: Feng, Duanyu, et al.
Veröffentlicht: (2023)
von: Feng, Duanyu, et al.
Veröffentlicht: (2023)
KaSA: Knowledge-Aware Singular-Value Adaptation of Large Language Models
von: Wang, Fan, et al.
Veröffentlicht: (2024)
von: Wang, Fan, et al.
Veröffentlicht: (2024)
The Structure of Relation Decoding Linear Operators in Large Language Models
von: Christ, Miranda Anna, et al.
Veröffentlicht: (2025)
von: Christ, Miranda Anna, et al.
Veröffentlicht: (2025)
CORE: Comprehensive Ontological Relation Evaluation for Large Language Models
von: Dwivedi, Satyam, et al.
Veröffentlicht: (2026)
von: Dwivedi, Satyam, et al.
Veröffentlicht: (2026)
Large Language Models as Markov Chains
von: Zekri, Oussama, et al.
Veröffentlicht: (2024)
von: Zekri, Oussama, et al.
Veröffentlicht: (2024)
Mitigating Biases for Instruction-following Language Models via Bias Neurons Elimination
von: Yang, Nakyeong, et al.
Veröffentlicht: (2023)
von: Yang, Nakyeong, et al.
Veröffentlicht: (2023)
WKVQuant: Quantizing Weight and Key/Value Cache for Large Language Models Gains More
von: Yue, Yuxuan, et al.
Veröffentlicht: (2024)
von: Yue, Yuxuan, et al.
Veröffentlicht: (2024)
Do Language Models Exhibit the Same Cognitive Biases in Problem Solving as Human Learners?
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
von: Opedal, Andreas, et al.
Veröffentlicht: (2024)
Subtle Biases Need Subtler Measures: Dual Metrics for Evaluating Representative and Affinity Bias in Large Language Models
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
von: Kumar, Abhishek, et al.
Veröffentlicht: (2024)
Laissez-Faire Harms: Algorithmic Biases in Generative Language Models
von: Shieh, Evan, et al.
Veröffentlicht: (2024)
von: Shieh, Evan, et al.
Veröffentlicht: (2024)
RE-Adapt: Reverse Engineered Adaptation of Large Language Models
von: Fleshman, William, et al.
Veröffentlicht: (2024)
von: Fleshman, William, et al.
Veröffentlicht: (2024)
A Taxonomy of Stereotype Content in Large Language Models
von: Nicolas, Gandalf, et al.
Veröffentlicht: (2024)
von: Nicolas, Gandalf, et al.
Veröffentlicht: (2024)
Interpreting the Repeated Token Phenomenon in Large Language Models
von: Yona, Itay, et al.
Veröffentlicht: (2025)
von: Yona, Itay, et al.
Veröffentlicht: (2025)
Meanings and Feelings of Large Language Models: Observability of Latent States in Generative AI
von: Liu, Tian Yu, et al.
Veröffentlicht: (2024)
von: Liu, Tian Yu, et al.
Veröffentlicht: (2024)
Heterogeneous Value Alignment Evaluation for Large Language Models
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2023)
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2023)
A Relative-Budget Theory for Reinforcement Learning with Verifiable Rewards in Large Language Model Reasoning
von: Wachi, Akifumi, et al.
Veröffentlicht: (2026)
von: Wachi, Akifumi, et al.
Veröffentlicht: (2026)
Leveraging Group Relative Policy Optimization to Advance Large Language Models in Traditional Chinese Medicine
von: Xie, Jiacheng, et al.
Veröffentlicht: (2025)
von: Xie, Jiacheng, et al.
Veröffentlicht: (2025)
zrLLM: Zero-Shot Relational Learning on Temporal Knowledge Graphs with Large Language Models
von: Ding, Zifeng, et al.
Veröffentlicht: (2023)
von: Ding, Zifeng, et al.
Veröffentlicht: (2023)
Value Augmented Sampling for Language Model Alignment and Personalization
von: Han, Seungwook, et al.
Veröffentlicht: (2024)
von: Han, Seungwook, et al.
Veröffentlicht: (2024)
Value-Aware Numerical Representations for Transformer Language Models
von: Dutulescu, Andreea, et al.
Veröffentlicht: (2026)
von: Dutulescu, Andreea, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Large Language Models are Biased Reinforcement Learners
von: Hayes, William M., et al.
Veröffentlicht: (2024) -
LogProber: Disentangling confidence from contamination in LLM responses
von: Yax, Nicolas, et al.
Veröffentlicht: (2024) -
PhyloLM : Inferring the Phylogeny of Large Language Models and Predicting their Performances in Benchmarks
von: Yax, Nicolas, et al.
Veröffentlicht: (2024) -
Large Language Models are Geographically Biased
von: Manvi, Rohin, et al.
Veröffentlicht: (2024) -
Active Preference Learning for Large Language Models
von: Muldrew, William, et al.
Veröffentlicht: (2024)