Self-Alignment: Improving Alignment of Cultural Values in LLMs via In-Context Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Choenni, Rochelle, Shutova, Ekaterina |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Echoes of Multilinguality: Tracing Cultural Value Shifts during LM Fine-tuning
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
How do languages influence each other? Studying cross-lingual data sharing during LM fine-tuning
di: Choenni, Rochelle, et al.
Pubblicazione: (2023)
di: Choenni, Rochelle, et al.
Pubblicazione: (2023)
Metaphor Understanding Challenge Dataset for LLMs
di: Tong, Xiaoyu, et al.
Pubblicazione: (2024)
di: Tong, Xiaoyu, et al.
Pubblicazione: (2024)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
di: Choenni, Rochelle, et al.
Pubblicazione: (2024)
Best-of-L: Cross-Lingual Reward Modeling for Mathematical Reasoning
di: Rajaee, Sara, et al.
Pubblicazione: (2025)
di: Rajaee, Sara, et al.
Pubblicazione: (2025)
M-Wanda: Improving One-Shot Pruning for Multilingual LLMs
di: Choenni, Rochelle, et al.
Pubblicazione: (2025)
di: Choenni, Rochelle, et al.
Pubblicazione: (2025)
Beyond Words: Exploring Cultural Value Sensitivity in Multimodal Models
di: Yadav, Srishti, et al.
Pubblicazione: (2025)
di: Yadav, Srishti, et al.
Pubblicazione: (2025)
Finding Culture-Sensitive Neurons in Vision-Language Models
di: Zhao, Xiutian, et al.
Pubblicazione: (2025)
di: Zhao, Xiutian, et al.
Pubblicazione: (2025)
Induction Heads as an Essential Mechanism for Pattern Matching in In-context Learning
di: Crosbie, Joy, et al.
Pubblicazione: (2024)
di: Crosbie, Joy, et al.
Pubblicazione: (2024)
Local Contrastive Editing of Gender Stereotypes
di: Lutz, Marlene, et al.
Pubblicazione: (2024)
di: Lutz, Marlene, et al.
Pubblicazione: (2024)
Are LLMs classical or nonmonotonic reasoners? Lessons from generics
di: Leidinger, Alina, et al.
Pubblicazione: (2024)
di: Leidinger, Alina, et al.
Pubblicazione: (2024)
PICACO: Pluralistic In-Context Value Alignment of LLMs via Total Correlation Optimization
di: Jiang, Han, et al.
Pubblicazione: (2025)
di: Jiang, Han, et al.
Pubblicazione: (2025)
A framework for annotating and modelling intentions behind metaphor use
di: Michelli, Gianluca, et al.
Pubblicazione: (2024)
di: Michelli, Gianluca, et al.
Pubblicazione: (2024)
Density Matrices for Metaphor Understanding
di: Owers, Jay, et al.
Pubblicazione: (2024)
di: Owers, Jay, et al.
Pubblicazione: (2024)
Learning New Tasks from a Few Examples with Soft-Label Prototypes
di: Singh, Avyav Kumar, et al.
Pubblicazione: (2022)
di: Singh, Avyav Kumar, et al.
Pubblicazione: (2022)
Yesterday's News: Benchmarking Multi-Dimensional Out-of-Distribution Generalization of Misinformation Detection Models
di: Verhoeven, Ivo, et al.
Pubblicazione: (2024)
di: Verhoeven, Ivo, et al.
Pubblicazione: (2024)
I Am Aligned, But With Whom? MENA Values Benchmark for Evaluating Cultural Alignment and Multilingual Bias in LLMs
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2025)
di: Zahraei, Pardis Sadat, et al.
Pubblicazione: (2025)
Survey-to-Behavior: Downstream Alignment of Human Values in LLMs via Survey Questions
di: Nie, Shangrui, et al.
Pubblicazione: (2025)
di: Nie, Shangrui, et al.
Pubblicazione: (2025)
Flames: Benchmarking Value Alignment of LLMs in Chinese
di: Huang, Kexin, et al.
Pubblicazione: (2023)
di: Huang, Kexin, et al.
Pubblicazione: (2023)
Improving the Distributional Alignment of LLMs using Supervision
di: Kambhatla, Gauri, et al.
Pubblicazione: (2025)
di: Kambhatla, Gauri, et al.
Pubblicazione: (2025)
Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
di: Zhang, Xiaoying, et al.
Pubblicazione: (2024)
Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs
di: Anantaprayoon, Panatchakorn, et al.
Pubblicazione: (2026)
di: Anantaprayoon, Panatchakorn, et al.
Pubblicazione: (2026)
Self-Pluralising Culture Alignment for Large Language Models
di: Xu, Shaoyang, et al.
Pubblicazione: (2024)
di: Xu, Shaoyang, et al.
Pubblicazione: (2024)
Assessing Socio-Cultural Alignment and Technical Safety of Sovereign LLMs
di: Chae, Kyubyung, et al.
Pubblicazione: (2025)
di: Chae, Kyubyung, et al.
Pubblicazione: (2025)
To Generate or Discriminate? Methodological Considerations for Measuring Cultural Alignment in LLMs
di: Pandey, Saurabh Kumar, et al.
Pubblicazione: (2026)
di: Pandey, Saurabh Kumar, et al.
Pubblicazione: (2026)
Speak in Context: Multilingual ASR with Speech Context Alignment via Contrastive Learning
di: Zhang, Yuchen, et al.
Pubblicazione: (2026)
di: Zhang, Yuchen, et al.
Pubblicazione: (2026)
Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values?
di: Shen, Hua, et al.
Pubblicazione: (2025)
di: Shen, Hua, et al.
Pubblicazione: (2025)
Contextual Moral Value Alignment Through Context-Based Aggregation
di: Dognin, Pierre, et al.
Pubblicazione: (2024)
di: Dognin, Pierre, et al.
Pubblicazione: (2024)
Step-On-Feet Tuning: Scaling Self-Alignment of LLMs via Bootstrapping
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
Context Misleads LLMs: The Role of Context Filtering in Maintaining Safe Alignment of LLMs
di: Kim, Jinhwa, et al.
Pubblicazione: (2025)
di: Kim, Jinhwa, et al.
Pubblicazione: (2025)
Cultural Palette: Pluralising Culture Alignment via Multi-agent Palette
di: Yuan, Jiahao, et al.
Pubblicazione: (2024)
di: Yuan, Jiahao, et al.
Pubblicazione: (2024)
ALI-Agent: Assessing LLMs' Alignment with Human Values via Agent-based Evaluation
di: Zheng, Jingnan, et al.
Pubblicazione: (2024)
di: Zheng, Jingnan, et al.
Pubblicazione: (2024)
Can LLMs Express Personality Across Cultures? Introducing CulturalPersonas for Evaluating Trait Alignment
di: Dey, Priyanka, et al.
Pubblicazione: (2025)
di: Dey, Priyanka, et al.
Pubblicazione: (2025)
Improving Alignment in LVLMs with Debiased Self-Judgment
di: Yang, Sihan, et al.
Pubblicazione: (2025)
di: Yang, Sihan, et al.
Pubblicazione: (2025)
Long-Short Alignment for Effective Long-Context Modeling in LLMs
di: Du, Tianqi, et al.
Pubblicazione: (2025)
di: Du, Tianqi, et al.
Pubblicazione: (2025)
Alignment at Pre-training! Towards Native Alignment for Arabic LLMs
di: Liang, Juhao, et al.
Pubblicazione: (2024)
di: Liang, Juhao, et al.
Pubblicazione: (2024)
ValuesRAG: Enhancing Cultural Alignment Through Retrieval-Augmented Contextual Learning
di: Seo, Wonduk, et al.
Pubblicazione: (2025)
di: Seo, Wonduk, et al.
Pubblicazione: (2025)
PluralLLM: Pluralistic Alignment in LLMs via Federated Learning
di: Srewa, Mahmoud, et al.
Pubblicazione: (2025)
di: Srewa, Mahmoud, et al.
Pubblicazione: (2025)
Evaluating and Improving Cultural Awareness of Reward Models for LLM Alignment
di: Zhang, Hongbin, et al.
Pubblicazione: (2025)
di: Zhang, Hongbin, et al.
Pubblicazione: (2025)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
di: Vo, Truong, et al.
Pubblicazione: (2025)
di: Vo, Truong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
The Echoes of Multilinguality: Tracing Cultural Value Shifts during LM Fine-tuning
di: Choenni, Rochelle, et al.
Pubblicazione: (2024) -
How do languages influence each other? Studying cross-lingual data sharing during LM fine-tuning
di: Choenni, Rochelle, et al.
Pubblicazione: (2023) -
Metaphor Understanding Challenge Dataset for LLMs
di: Tong, Xiaoyu, et al.
Pubblicazione: (2024) -
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
di: Choenni, Rochelle, et al.
Pubblicazione: (2024) -
Best-of-L: Cross-Lingual Reward Modeling for Mathematical Reasoning
di: Rajaee, Sara, et al.
Pubblicazione: (2025)