How Value Induction Reshapes LLM Behaviour
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Arora, Arnav, Schluter, Natalie, Metcalf, Katherine, ter Hoeve, Maartje |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On the Way to LLM Personalization: Learning to Remember User Conversations
von: Magister, Lucie Charlotte, et al.
Veröffentlicht: (2024)
von: Magister, Lucie Charlotte, et al.
Veröffentlicht: (2024)
Multilingual Knowledge Transfer under Data Constraints via Lexical Interventions
von: Sedova, Anastasiia, et al.
Veröffentlicht: (2026)
von: Sedova, Anastasiia, et al.
Veröffentlicht: (2026)
GrammaMT: Improving Machine Translation with Grammar-Informed In-Context Learning
von: Ramos, Rita, et al.
Veröffentlicht: (2024)
von: Ramos, Rita, et al.
Veröffentlicht: (2024)
Training Bilingual LMs with Data Constraints in the Targeted Language
von: Seto, Skyler, et al.
Veröffentlicht: (2024)
von: Seto, Skyler, et al.
Veröffentlicht: (2024)
Discriminating Form and Meaning in Multilingual Models with Minimal-Pair ABX Tasks
von: de Seyssel, Maureen, et al.
Veröffentlicht: (2025)
von: de Seyssel, Maureen, et al.
Veröffentlicht: (2025)
Analyzing the Effect of Linguistic Similarity on Cross-Lingual Transfer: Tasks and Experimental Setups Matter
von: Blaschke, Verena, et al.
Veröffentlicht: (2025)
von: Blaschke, Verena, et al.
Veröffentlicht: (2025)
Assessing the Role of Data Quality in Training Bilingual Language Models
von: Seto, Skyler, et al.
Veröffentlicht: (2025)
von: Seto, Skyler, et al.
Veröffentlicht: (2025)
mRAKL: Multilingual Retrieval-Augmented Knowledge Graph Construction for Low-Resourced Languages
von: Nigatu, Hellina Hailu, et al.
Veröffentlicht: (2025)
von: Nigatu, Hellina Hailu, et al.
Veröffentlicht: (2025)
On the Limited Generalization Capability of the Implicit Reward Model Induced by Direct Preference Optimization
von: Lin, Yong, et al.
Veröffentlicht: (2024)
von: Lin, Yong, et al.
Veröffentlicht: (2024)
Analyzing Dialectical Biases in LLMs for Knowledge and Reasoning Benchmarks
von: Pan, Eileen, et al.
Veröffentlicht: (2025)
von: Pan, Eileen, et al.
Veröffentlicht: (2025)
Probing Pre-Trained Language Models for Cross-Cultural Differences in Values
von: Arora, Arnav, et al.
Veröffentlicht: (2022)
von: Arora, Arnav, et al.
Veröffentlicht: (2022)
Beyond the Battlefield: Framing Analysis of Media Coverage in Conflict Reporting
von: Kaur, Avneet, et al.
Veröffentlicht: (2025)
von: Kaur, Avneet, et al.
Veröffentlicht: (2025)
Presumed Cultural Identity: How Names Shape LLM Responses
von: Pawar, Siddhesh, et al.
Veröffentlicht: (2025)
von: Pawar, Siddhesh, et al.
Veröffentlicht: (2025)
Revealing Fine-Grained Values and Opinions in Large Language Models
von: Wright, Dustin, et al.
Veröffentlicht: (2024)
von: Wright, Dustin, et al.
Veröffentlicht: (2024)
Aligning LLMs by Predicting Preferences from User Writing Samples
von: Aroca-Ouellette, Stéphane, et al.
Veröffentlicht: (2025)
von: Aroca-Ouellette, Stéphane, et al.
Veröffentlicht: (2025)
Why Should This Article Be Deleted? Transparent Stance Detection in Multilingual Wikipedia Editor Discussions
von: Kaffee, Lucie-Aimée, et al.
Veröffentlicht: (2023)
von: Kaffee, Lucie-Aimée, et al.
Veröffentlicht: (2023)
Pitch Accent Detection improves Pretrained Automatic Speech Recognition
von: Sasu, David, et al.
Veröffentlicht: (2025)
von: Sasu, David, et al.
Veröffentlicht: (2025)
Overview of the 2023 ICON Shared Task on Gendered Abuse Detection in Indic Languages
von: Vaidya, Aatman, et al.
Veröffentlicht: (2024)
von: Vaidya, Aatman, et al.
Veröffentlicht: (2024)
The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas
von: Wu, Ya, et al.
Veröffentlicht: (2025)
von: Wu, Ya, et al.
Veröffentlicht: (2025)
The Role of Prosody in Spoken Question Answering
von: Chi, Jie, et al.
Veröffentlicht: (2025)
von: Chi, Jie, et al.
Veröffentlicht: (2025)
How Lexical is Bilingual Lexicon Induction?
von: Kohli, Harsh, et al.
Veröffentlicht: (2024)
von: Kohli, Harsh, et al.
Veröffentlicht: (2024)
A Multilingual, Large-Scale Study of the Interplay between LLM Safeguards, Personalisation, and Disinformation
von: Leite, João A., et al.
Veröffentlicht: (2025)
von: Leite, João A., et al.
Veröffentlicht: (2025)
Fingerprinting New Physics with Effective Field Theories
von: ter Hoeve, Jaco
Veröffentlicht: (2025)
von: ter Hoeve, Jaco
Veröffentlicht: (2025)
Multi-Modal Framing Analysis of News
von: Arora, Arnav, et al.
Veröffentlicht: (2025)
von: Arora, Arnav, et al.
Veröffentlicht: (2025)
Surgical Feature-Space Decomposition of LLMs: Why, When and How?
von: Chavan, Arnav, et al.
Veröffentlicht: (2024)
von: Chavan, Arnav, et al.
Veröffentlicht: (2024)
Is Grep All You Need? How Agent Harnesses Reshape Agentic Search
von: Sen, Sahil, et al.
Veröffentlicht: (2026)
von: Sen, Sahil, et al.
Veröffentlicht: (2026)
Specializing Large Language Models to Simulate Survey Response Distributions for Global Populations
von: Cao, Yong, et al.
Veröffentlicht: (2025)
von: Cao, Yong, et al.
Veröffentlicht: (2025)
HORIZON: A Benchmark for In-the-wild User Behaviour Modeling
von: Goel, Arnav, et al.
Veröffentlicht: (2026)
von: Goel, Arnav, et al.
Veröffentlicht: (2026)
In the LLM era, Word Sense Induction remains unsolved
von: Mosolova, Anna, et al.
Veröffentlicht: (2026)
von: Mosolova, Anna, et al.
Veröffentlicht: (2026)
Dialog Flow Induction for Constrainable LLM-Based Chatbots
von: Agrawal, Stuti, et al.
Veröffentlicht: (2024)
von: Agrawal, Stuti, et al.
Veröffentlicht: (2024)
Positional Fragility in LLMs: How Offset Effects Reshape Our Understanding of Memorization Risks
von: Xu, Yixuan, et al.
Veröffentlicht: (2025)
von: Xu, Yixuan, et al.
Veröffentlicht: (2025)
Scaling Laws for Mixture Pretraining Under Data Constraints
von: Sedova, Anastasiia, et al.
Veröffentlicht: (2026)
von: Sedova, Anastasiia, et al.
Veröffentlicht: (2026)
Evaluation Drift in LLM Personality Induction: Are We Moving the Goalpost?
von: Rajput, Prateek, et al.
Veröffentlicht: (2026)
von: Rajput, Prateek, et al.
Veröffentlicht: (2026)
From MOOC to MAIC: Reshaping Online Teaching and Learning through LLM-driven Agents
von: Yu, Jifan, et al.
Veröffentlicht: (2024)
von: Yu, Jifan, et al.
Veröffentlicht: (2024)
Contextual Drag: How Errors in the Context Affect LLM Reasoning
von: Cheng, Yun, et al.
Veröffentlicht: (2026)
von: Cheng, Yun, et al.
Veröffentlicht: (2026)
Narrative-to-Scene Generation: An LLM-Driven Pipeline for 2D Game Environments
von: Chen, Yi-Chun, et al.
Veröffentlicht: (2025)
von: Chen, Yi-Chun, et al.
Veröffentlicht: (2025)
RLSF: Fine-tuning LLMs via Symbolic Feedback
von: Jha, Piyush, et al.
Veröffentlicht: (2024)
von: Jha, Piyush, et al.
Veröffentlicht: (2024)
Same Model, Different Weakness: How Language and Modality Reshape the Jailbreak Attack Surface in Frontier MLLMs
von: Ford, Casey, et al.
Veröffentlicht: (2026)
von: Ford, Casey, et al.
Veröffentlicht: (2026)
ProtocolLLM: RTL Benchmark for SystemVerilog Generation of Communication Protocols
von: Sheth, Arnav, et al.
Veröffentlicht: (2025)
von: Sheth, Arnav, et al.
Veröffentlicht: (2025)
Designing Role Vectors to Improve LLM Inference Behaviour
von: Potertì, Daniele, et al.
Veröffentlicht: (2025)
von: Potertì, Daniele, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
On the Way to LLM Personalization: Learning to Remember User Conversations
von: Magister, Lucie Charlotte, et al.
Veröffentlicht: (2024) -
Multilingual Knowledge Transfer under Data Constraints via Lexical Interventions
von: Sedova, Anastasiia, et al.
Veröffentlicht: (2026) -
GrammaMT: Improving Machine Translation with Grammar-Informed In-Context Learning
von: Ramos, Rita, et al.
Veröffentlicht: (2024) -
Training Bilingual LMs with Data Constraints in the Targeted Language
von: Seto, Skyler, et al.
Veröffentlicht: (2024) -
Discriminating Form and Meaning in Multilingual Models with Minimal-Pair ABX Tasks
von: de Seyssel, Maureen, et al.
Veröffentlicht: (2025)