An Investigation into Value Misalignment in LLM-Generated Texts for Cultural Heritage
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bu, Fan, Wang, Zheng, Wang, Siyi, Liu, Ziyao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LLM Hallucinations in Practical Code Generation: Phenomena, Mechanism, and Mitigation
von: Zhang, Ziyao, et al.
Veröffentlicht: (2024)
von: Zhang, Ziyao, et al.
Veröffentlicht: (2024)
Lost in Translation: Latent Concept Misalignment in Text-to-Image Diffusion Models
von: Zhao, Juntu, et al.
Veröffentlicht: (2024)
von: Zhao, Juntu, et al.
Veröffentlicht: (2024)
Preemptive Detection and Correction of Misaligned Actions in LLM Agents
von: Fang, Haishuo, et al.
Veröffentlicht: (2024)
von: Fang, Haishuo, et al.
Veröffentlicht: (2024)
Soundwave: Less is More for Speech-Text Alignment in LLMs
von: Zhang, Yuhao, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2025)
Emergent Misalignment is Easy, Narrow Misalignment is Hard
von: Soligo, Anna, et al.
Veröffentlicht: (2026)
von: Soligo, Anna, et al.
Veröffentlicht: (2026)
Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement
von: Lee, Daeun, et al.
Veröffentlicht: (2024)
von: Lee, Daeun, et al.
Veröffentlicht: (2024)
The African Woman is Rhythmic and Soulful: An Investigation of Implicit Biases in LLM Open-ended Text Generation
von: Lim, Serene, et al.
Veröffentlicht: (2024)
von: Lim, Serene, et al.
Veröffentlicht: (2024)
Fusing Bidirectional Chains of Thought and Reward Mechanisms A Method for Enhancing Question-Answering Capabilities of Large Language Models for Chinese Intangible Cultural Heritage
von: Liu, Ruilin, et al.
Veröffentlicht: (2025)
von: Liu, Ruilin, et al.
Veröffentlicht: (2025)
SHIELD: Evaluation and Defense Strategies for Copyright Compliance in LLM Text Generation
von: Liu, Xiaoze, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoze, et al.
Veröffentlicht: (2024)
ValuesRAG: Enhancing Cultural Alignment Through Retrieval-Augmented Contextual Learning
von: Seo, Wonduk, et al.
Veröffentlicht: (2025)
von: Seo, Wonduk, et al.
Veröffentlicht: (2025)
Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates
von: Schimmenti, Andrea, et al.
Veröffentlicht: (2025)
von: Schimmenti, Andrea, et al.
Veröffentlicht: (2025)
LLMs Deceive Unintentionally: Emergent Misalignment in Dishonesty from Misaligned Samples to Biased Human-AI Interactions
von: Hu, Xuhao, et al.
Veröffentlicht: (2025)
von: Hu, Xuhao, et al.
Veröffentlicht: (2025)
GenQuest: An LLM-based Text Adventure Game for Language Learners
von: Wang, Qiao, et al.
Veröffentlicht: (2025)
von: Wang, Qiao, et al.
Veröffentlicht: (2025)
Towards Geo-Culturally Grounded LLM Generations
von: Lertvittayakumjorn, Piyawat, et al.
Veröffentlicht: (2025)
von: Lertvittayakumjorn, Piyawat, et al.
Veröffentlicht: (2025)
Estimating Causal Effects of Text Interventions Leveraging LLMs
von: Guo, Siyi, et al.
Veröffentlicht: (2024)
von: Guo, Siyi, et al.
Veröffentlicht: (2024)
Variation is the Key: A Variation-Based Framework for LLM-Generated Text Detection
von: Li, Xuecong, et al.
Veröffentlicht: (2026)
von: Li, Xuecong, et al.
Veröffentlicht: (2026)
Bringing Value Models Back: Generative Critics for Value Modeling in LLM Reinforcement Learning
von: Shan, Zikang, et al.
Veröffentlicht: (2026)
von: Shan, Zikang, et al.
Veröffentlicht: (2026)
Mitigating Misalignment Contagion by Steering with Implicit Traits
von: Chang, Maria, et al.
Veröffentlicht: (2026)
von: Chang, Maria, et al.
Veröffentlicht: (2026)
Raising the Bar: Investigating the Values of Large Language Models via Generative Evolving Testing
von: Jiang, Han, et al.
Veröffentlicht: (2024)
von: Jiang, Han, et al.
Veröffentlicht: (2024)
Generation-driven Contrastive Self-training for Zero-shot Text Classification with Instruction-following LLM
von: Zhang, Ruohong, et al.
Veröffentlicht: (2023)
von: Zhang, Ruohong, et al.
Veröffentlicht: (2023)
Generative AI in Heritage Practice: Improving the Accessibility of Heritage Guidance
von: Witte, Jessica, et al.
Veröffentlicht: (2025)
von: Witte, Jessica, et al.
Veröffentlicht: (2025)
Graphs of Research: Citation Evolution Graphs as Supervision for Research Idea Generation
von: Gao, Songyang, et al.
Veröffentlicht: (2026)
von: Gao, Songyang, et al.
Veröffentlicht: (2026)
Investigating LLM Applications in E-Commerce
von: Palen-Michel, Chester, et al.
Veröffentlicht: (2024)
von: Palen-Michel, Chester, et al.
Veröffentlicht: (2024)
Is Your Paper Being Reviewed by an LLM? Investigating AI Text Detectability in Peer Review
von: Yu, Sungduk, et al.
Veröffentlicht: (2024)
von: Yu, Sungduk, et al.
Veröffentlicht: (2024)
CLAVE: An Adaptive Framework for Evaluating Values of LLM Generated Responses
von: Yao, Jing, et al.
Veröffentlicht: (2024)
von: Yao, Jing, et al.
Veröffentlicht: (2024)
Zero-Shot Detection of LLM-Generated Text via Implicit Reward Model
von: Liu, Runheng, et al.
Veröffentlicht: (2026)
von: Liu, Runheng, et al.
Veröffentlicht: (2026)
Generative Value Conflicts Reveal LLM Priorities
von: Liu, Andy, et al.
Veröffentlicht: (2025)
von: Liu, Andy, et al.
Veröffentlicht: (2025)
RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
LLMs and Cultural Values: the Impact of Prompt Language and Explicit Cultural Framing
von: Bulté, Bram, et al.
Veröffentlicht: (2025)
von: Bulté, Bram, et al.
Veröffentlicht: (2025)
Distributional Open-Ended Evaluation of LLM Cultural Value Alignment Based on Value Codebook
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2026)
von: Lee, Jaehyeok, et al.
Veröffentlicht: (2026)
Explainability-Based Token Replacement on LLM-Generated Text
von: Mohammadi, Hadi, et al.
Veröffentlicht: (2025)
von: Mohammadi, Hadi, et al.
Veröffentlicht: (2025)
Fairness or Fluency? An Investigation into Language Bias of Pairwise LLM-as-a-Judge
von: Zhou, Xiaolin, et al.
Veröffentlicht: (2026)
von: Zhou, Xiaolin, et al.
Veröffentlicht: (2026)
On the Detectability of LLM-Generated Text: What Exactly Is LLM-Generated Text?
von: Geng, Mingmeng, et al.
Veröffentlicht: (2025)
von: Geng, Mingmeng, et al.
Veröffentlicht: (2025)
Next-Generation Database Interfaces: A Survey of LLM-based Text-to-SQL
von: Hong, Zijin, et al.
Veröffentlicht: (2024)
von: Hong, Zijin, et al.
Veröffentlicht: (2024)
Reviving Cultural Heritage: A Novel Approach for Comprehensive Historical Document Restoration
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes
von: Masoudian, Shahed, et al.
Veröffentlicht: (2025)
von: Masoudian, Shahed, et al.
Veröffentlicht: (2025)
PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models
von: An, Jiyuan, et al.
Veröffentlicht: (2026)
von: An, Jiyuan, et al.
Veröffentlicht: (2026)
MRScore: Evaluating Radiology Report Generation with LLM-based Reward System
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
von: Liu, Yunyi, et al.
Veröffentlicht: (2024)
Alleviating Choice Supportive Bias in LLM with Reasoning Dependency Generation
von: Zhuang, Nan, et al.
Veröffentlicht: (2025)
von: Zhuang, Nan, et al.
Veröffentlicht: (2025)
Rethinking DPO: The Role of Rejected Responses in Preference Misalignment
von: Cho, Jay Hyeon, et al.
Veröffentlicht: (2025)
von: Cho, Jay Hyeon, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
LLM Hallucinations in Practical Code Generation: Phenomena, Mechanism, and Mitigation
von: Zhang, Ziyao, et al.
Veröffentlicht: (2024) -
Lost in Translation: Latent Concept Misalignment in Text-to-Image Diffusion Models
von: Zhao, Juntu, et al.
Veröffentlicht: (2024) -
Preemptive Detection and Correction of Misaligned Actions in LLM Agents
von: Fang, Haishuo, et al.
Veröffentlicht: (2024) -
Soundwave: Less is More for Speech-Text Alignment in LLMs
von: Zhang, Yuhao, et al.
Veröffentlicht: (2025) -
Emergent Misalignment is Easy, Narrow Misalignment is Hard
von: Soligo, Anna, et al.
Veröffentlicht: (2026)