Exploring the Impact of Instruction-Tuning on LLM's Susceptibility to Misinformation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Kyubeen, Jang, Junseo, Kim, Hongjin, Jeong, Geunyeong, Kim, Harksoo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
STEAM: A Semantic-Level Knowledge Editing Framework for Large Language Models
von: Jeong, Geunyeong, et al.
Veröffentlicht: (2025)
von: Jeong, Geunyeong, et al.
Veröffentlicht: (2025)
Can Large Language Models Differentiate Harmful from Argumentative Essays? Steps Toward Ethical Essay Scoring
von: Kim, Hongjin, et al.
Veröffentlicht: (2026)
von: Kim, Hongjin, et al.
Veröffentlicht: (2026)
Generation-Based and Emotion-Reflected Memory Update: Creating the KEEM Dataset for Better Long-Term Conversation
von: Kang, Jeonghyun, et al.
Veröffentlicht: (2026)
von: Kang, Jeonghyun, et al.
Veröffentlicht: (2026)
Breaking the Pre-Sampling Barrier: Activation-Informed Difficulty-Aware Self-Consistency
von: Yoon, Taewoong, et al.
Veröffentlicht: (2026)
von: Yoon, Taewoong, et al.
Veröffentlicht: (2026)
ReviewScore: Misinformed Peer Review Detection with Large Language Models
von: Ryu, Hyun, et al.
Veröffentlicht: (2025)
von: Ryu, Hyun, et al.
Veröffentlicht: (2025)
KAIO: A Collection of More Challenging Korean Questions
von: Lee, Nahyun, et al.
Veröffentlicht: (2025)
von: Lee, Nahyun, et al.
Veröffentlicht: (2025)
Instructive Decoding: Instruction-Tuned Large Language Models are Self-Refiner from Noisy Instructions
von: Kim, Taehyeon, et al.
Veröffentlicht: (2023)
von: Kim, Taehyeon, et al.
Veröffentlicht: (2023)
Generative Representational Instruction Tuning
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
von: Muennighoff, Niklas, et al.
Veröffentlicht: (2024)
Overstating Attitudes, Ignoring Networks: LLM Biases in Simulating Misinformation Susceptibility
von: Choi, Eun Cheol, et al.
Veröffentlicht: (2026)
von: Choi, Eun Cheol, et al.
Veröffentlicht: (2026)
Evaluating the Simulation of Human Personality-Driven Susceptibility to Misinformation with LLMs
von: Pratelli, Manuel, et al.
Veröffentlicht: (2025)
von: Pratelli, Manuel, et al.
Veröffentlicht: (2025)
Instruction Tuning with and without Context: Behavioral Shifts and Downstream Impact
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
Fast KVzip: Efficient and Accurate LLM Inference with Gated KV Eviction
von: Kim, Jang-Hyun, et al.
Veröffentlicht: (2026)
von: Kim, Jang-Hyun, et al.
Veröffentlicht: (2026)
Chain-of-Instructions: Compositional Instruction Tuning on Large Language Models
von: Hayati, Shirley Anugrah, et al.
Veröffentlicht: (2024)
von: Hayati, Shirley Anugrah, et al.
Veröffentlicht: (2024)
Explore the Potential of LLMs in Misinformation Detection: An Empirical Study
von: Chen, Mengyang, et al.
Veröffentlicht: (2023)
von: Chen, Mengyang, et al.
Veröffentlicht: (2023)
Multi-View Attention Multiple-Instance Learning Enhanced by LLM Reasoning for Cognitive Distortion Detection
von: Kim, Jun Seo, et al.
Veröffentlicht: (2025)
von: Kim, Jun Seo, et al.
Veröffentlicht: (2025)
Exploring Format Consistency for Instruction Tuning
von: Liang, Shihao, et al.
Veröffentlicht: (2023)
von: Liang, Shihao, et al.
Veröffentlicht: (2023)
Unraveling Misinformation Propagation in LLM Reasoning
von: Feng, Yiyang, et al.
Veröffentlicht: (2025)
von: Feng, Yiyang, et al.
Veröffentlicht: (2025)
KMMMU: Evaluation of Massive Multi-discipline Multimodal Understanding in Korean Language and Context
von: Lee, Nahyun, et al.
Veröffentlicht: (2026)
von: Lee, Nahyun, et al.
Veröffentlicht: (2026)
A Survey on Data Selection for LLM Instruction Tuning
von: Zhang, Bolin, et al.
Veröffentlicht: (2024)
von: Zhang, Bolin, et al.
Veröffentlicht: (2024)
LLM-Detector: Improving AI-Generated Chinese Text Detection with Open-Source LLM Instruction Tuning
von: Wang, Rongsheng, et al.
Veröffentlicht: (2024)
von: Wang, Rongsheng, et al.
Veröffentlicht: (2024)
A Cognitively Grounded Bayesian Framework for Misinformation Susceptibility
von: Madhyastha, Pranava
Veröffentlicht: (2026)
von: Madhyastha, Pranava
Veröffentlicht: (2026)
CoachLM: Automatic Instruction Revisions Improve the Data Quality in LLM Instruction Tuning
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
von: Liu, Yilun, et al.
Veröffentlicht: (2023)
Merging Triggers, Breaking Backdoors: Defensive Poisoning for Instruction-Tuned Language Models
von: Kim, San, et al.
Veröffentlicht: (2026)
von: Kim, San, et al.
Veröffentlicht: (2026)
Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility
von: Borah, Angana, et al.
Veröffentlicht: (2026)
von: Borah, Angana, et al.
Veröffentlicht: (2026)
Importance-Aware Data Selection for Efficient LLM Instruction Tuning
von: Jiang, Tingyu, et al.
Veröffentlicht: (2025)
von: Jiang, Tingyu, et al.
Veröffentlicht: (2025)
Know the Unknown: An Uncertainty-Sensitive Method for LLM Instruction Tuning
von: Li, Jiaqi, et al.
Veröffentlicht: (2024)
von: Li, Jiaqi, et al.
Veröffentlicht: (2024)
Star-Agents: Automatic Data Optimization with LLM Agents for Instruction Tuning
von: Zhou, Hang, et al.
Veröffentlicht: (2024)
von: Zhou, Hang, et al.
Veröffentlicht: (2024)
ISD-Agent-Bench: A Comprehensive Benchmark for Evaluating LLM-based Instructional Design Agents
von: Jeon, YoungHoon, et al.
Veröffentlicht: (2026)
von: Jeon, YoungHoon, et al.
Veröffentlicht: (2026)
NUS-Emo at SemEval-2024 Task 3: Instruction-Tuning LLM for Multimodal Emotion-Cause Analysis in Conversations
von: Luo, Meng, et al.
Veröffentlicht: (2024)
von: Luo, Meng, et al.
Veröffentlicht: (2024)
Decoding Susceptibility: Modeling Misbelief to Misinformation Through a Computational Approach
von: Liu, Yanchen, et al.
Veröffentlicht: (2023)
von: Liu, Yanchen, et al.
Veröffentlicht: (2023)
Exploring Text Representations for Online Misinformation
von: Dogo, Martins Samuel
Veröffentlicht: (2024)
von: Dogo, Martins Samuel
Veröffentlicht: (2024)
Exploring Coding Spot: Understanding Parametric Contributions to LLM Coding Performance
von: Kim, Dongjun, et al.
Veröffentlicht: (2024)
von: Kim, Dongjun, et al.
Veröffentlicht: (2024)
Capability Instruction Tuning: A New Paradigm for Dynamic LLM Routing
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2025)
von: Zhang, Yi-Kai, et al.
Veröffentlicht: (2025)
Tool-MAD: A Multi-Agent Debate Framework for Fact Verification with Diverse Tool Augmentation and Adaptive Retrieval
von: Jeong, Seyeon, et al.
Veröffentlicht: (2026)
von: Jeong, Seyeon, et al.
Veröffentlicht: (2026)
Exploring the Impact of Occupational Personas on Domain-Specific QA
von: Kang, Eojin, et al.
Veröffentlicht: (2025)
von: Kang, Eojin, et al.
Veröffentlicht: (2025)
PVP: An Image Dataset for Personalized Visual Persuasion with Persuasion Strategies, Viewer Characteristics, and Persuasiveness Ratings
von: Kim, Junseo, et al.
Veröffentlicht: (2025)
von: Kim, Junseo, et al.
Veröffentlicht: (2025)
EXAONE 3.0 7.8B Instruction Tuned Language Model
von: An, Soyoung, et al.
Veröffentlicht: (2024)
von: An, Soyoung, et al.
Veröffentlicht: (2024)
Model Fusion through Bayesian Optimization in Language Model Fine-Tuning
von: Jang, Chaeyun, et al.
Veröffentlicht: (2024)
von: Jang, Chaeyun, et al.
Veröffentlicht: (2024)
Bayesian Multi-Task Transfer Learning for Soft Prompt Tuning
von: Lee, Haeju, et al.
Veröffentlicht: (2024)
von: Lee, Haeju, et al.
Veröffentlicht: (2024)
Instruction Following without Instruction Tuning
von: Hewitt, John, et al.
Veröffentlicht: (2024)
von: Hewitt, John, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
STEAM: A Semantic-Level Knowledge Editing Framework for Large Language Models
von: Jeong, Geunyeong, et al.
Veröffentlicht: (2025) -
Can Large Language Models Differentiate Harmful from Argumentative Essays? Steps Toward Ethical Essay Scoring
von: Kim, Hongjin, et al.
Veröffentlicht: (2026) -
Generation-Based and Emotion-Reflected Memory Update: Creating the KEEM Dataset for Better Long-Term Conversation
von: Kang, Jeonghyun, et al.
Veröffentlicht: (2026) -
Breaking the Pre-Sampling Barrier: Activation-Informed Difficulty-Aware Self-Consistency
von: Yoon, Taewoong, et al.
Veröffentlicht: (2026) -
ReviewScore: Misinformed Peer Review Detection with Large Language Models
von: Ryu, Hyun, et al.
Veröffentlicht: (2025)