DeFrame: Debiasing Large Language Models Against Framing Effects
Fuente:
arXiv
Salvato in:
| Autori principali: | Lim, Kahee, Kim, Soyeon, Whang, Steven Euijong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ERBench: An Entity-Relationship based Automatically Verifiable Hallucination Benchmark for Large Language Models
di: Oh, Jio, et al.
Pubblicazione: (2024)
di: Oh, Jio, et al.
Pubblicazione: (2024)
Classroom AI: Large Language Models as Grade-Specific Teachers
di: Oh, Jio, et al.
Pubblicazione: (2026)
di: Oh, Jio, et al.
Pubblicazione: (2026)
Harnessing Temporal Databases for Systematic Evaluation of Factual Time-Sensitive Question-Answering in Large Language Models
di: Kim, Soyeon, et al.
Pubblicazione: (2025)
di: Kim, Soyeon, et al.
Pubblicazione: (2025)
PFGuard: A Generative Framework with Privacy and Fairness Safeguards
di: Kim, Soyeon, et al.
Pubblicazione: (2024)
di: Kim, Soyeon, et al.
Pubblicazione: (2024)
DialectLLM: A Dialect-Aware Dialog[ue] Generation Framework Beyond Standard American English
di: Oh, Jio, et al.
Pubblicazione: (2026)
di: Oh, Jio, et al.
Pubblicazione: (2026)
Framing Effects in Independent-Agent Large Language Models: A Cross-Family Behavioral Analysis
di: Wang, Zice, et al.
Pubblicazione: (2026)
di: Wang, Zice, et al.
Pubblicazione: (2026)
GradMix: Gradient-based Selective Mixup for Robust Data Augmentation in Class-Incremental Learning
di: Kim, Minsu, et al.
Pubblicazione: (2025)
di: Kim, Minsu, et al.
Pubblicazione: (2025)
Evaluating Large Language Models on the Frame and Symbol Grounding Problems: A Zero-shot Benchmark
di: Oka, Shoko
Pubblicazione: (2025)
di: Oka, Shoko
Pubblicazione: (2025)
Mechanistic Interpretability of Socio-Political Frames in Language Models
di: Asghari, Hadi, et al.
Pubblicazione: (2025)
di: Asghari, Hadi, et al.
Pubblicazione: (2025)
Fair Class-Incremental Learning using Sample Weighting
di: Park, Jaeyoung, et al.
Pubblicazione: (2024)
di: Park, Jaeyoung, et al.
Pubblicazione: (2024)
UGID: Unified Graph Isomorphism for Debiasing Large Language Models
di: Ding, Zikang, et al.
Pubblicazione: (2026)
di: Ding, Zikang, et al.
Pubblicazione: (2026)
Causal-Guided Active Learning for Debiasing Large Language Models
di: Du, Li, et al.
Pubblicazione: (2024)
di: Du, Li, et al.
Pubblicazione: (2024)
A Scalable Pipeline for Estimating Verb Frame Frequencies Using Large Language Models
di: Morgan, Adam M., et al.
Pubblicazione: (2025)
di: Morgan, Adam M., et al.
Pubblicazione: (2025)
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
di: Lior, Gili, et al.
Pubblicazione: (2025)
di: Lior, Gili, et al.
Pubblicazione: (2025)
DirecT2V: Large Language Models are Frame-Level Directors for Zero-Shot Text-to-Video Generation
di: Hong, Susung, et al.
Pubblicazione: (2023)
di: Hong, Susung, et al.
Pubblicazione: (2023)
BiasLab: A Multilingual, Dual-Framing Framework for Robust Measurement of Output-Level Bias in Large Language Models
di: Guey, William, et al.
Pubblicazione: (2026)
di: Guey, William, et al.
Pubblicazione: (2026)
Framing Political Bias in Multilingual LLMs Across Pakistani Languages
di: Nadeem, Afrozah, et al.
Pubblicazione: (2025)
di: Nadeem, Afrozah, et al.
Pubblicazione: (2025)
How Prompts Move Language Model Behavior: Frames, Salience, and Construal as Semantic Control
di: Kim, Dongseok, et al.
Pubblicazione: (2025)
di: Kim, Dongseok, et al.
Pubblicazione: (2025)
Scaling Video-Language Models to 10K Frames via Hierarchical Differential Distillation
di: Cheng, Chuanqi, et al.
Pubblicazione: (2025)
di: Cheng, Chuanqi, et al.
Pubblicazione: (2025)
LLMs and Cultural Values: the Impact of Prompt Language and Explicit Cultural Framing
di: Bulté, Bram, et al.
Pubblicazione: (2025)
di: Bulté, Bram, et al.
Pubblicazione: (2025)
DeCAP: Context-Adaptive Prompt Generation for Debiasing Zero-shot Question Answering in Large Language Models
di: Bae, Suyoung, et al.
Pubblicazione: (2025)
di: Bae, Suyoung, et al.
Pubblicazione: (2025)
Debiasing Large Language Models via Adaptive Causal Prompting with Sketch-of-Thought
di: Li, Bowen, et al.
Pubblicazione: (2026)
di: Li, Bowen, et al.
Pubblicazione: (2026)
Information Gain-Guided Causal Intervention for Autonomous Debiasing Large Language Models
di: Sun, Zhouhao, et al.
Pubblicazione: (2025)
di: Sun, Zhouhao, et al.
Pubblicazione: (2025)
DataFrame QA: A Universal LLM Framework on DataFrame Question Answering Without Data Exposure
di: Ye, Junyi, et al.
Pubblicazione: (2024)
di: Ye, Junyi, et al.
Pubblicazione: (2024)
General Phrase Debiaser: Debiasing Masked Language Models at a Multi-Token Level
di: Shi, Bingkang, et al.
Pubblicazione: (2023)
di: Shi, Bingkang, et al.
Pubblicazione: (2023)
Bi-directional Bias Attribution: Debiasing Large Language Models without Modifying Prompts
di: Lin, Yujie, et al.
Pubblicazione: (2026)
di: Lin, Yujie, et al.
Pubblicazione: (2026)
Debiasing Large Language Models in Thai Political Stance Detection via Counterfactual Calibration
di: Sermsri, Kasidit, et al.
Pubblicazione: (2025)
di: Sermsri, Kasidit, et al.
Pubblicazione: (2025)
Team QUST at SemEval-2025 Task 10: Evaluating Large Language Models in Multiclass Multi-label Classification of News Entity Framing
di: Liu, Jiyan, et al.
Pubblicazione: (2025)
di: Liu, Jiyan, et al.
Pubblicazione: (2025)
FrameNet Semantic Role Classification by Analogy
di: Ngo, Van-Duy, et al.
Pubblicazione: (2026)
di: Ngo, Van-Duy, et al.
Pubblicazione: (2026)
Under Pressure: Emotional Framing Induces Measurable Behavioral Shifts and Structured Internal Geometry in Small Language Models
di: Usman, Rana Muhammad
Pubblicazione: (2026)
di: Usman, Rana Muhammad
Pubblicazione: (2026)
Self-Debiasing Large Language Models: Zero-Shot Recognition and Reduction of Stereotypes
di: Gallegos, Isabel O., et al.
Pubblicazione: (2024)
di: Gallegos, Isabel O., et al.
Pubblicazione: (2024)
Auto-Prompting with Retrieval Guidance for Frame Detection in Logistics
di: Duc, Do Minh, et al.
Pubblicazione: (2025)
di: Duc, Do Minh, et al.
Pubblicazione: (2025)
Creativity Has Left the Chat: The Price of Debiasing Language Models
di: Mohammadi, Behnam
Pubblicazione: (2024)
di: Mohammadi, Behnam
Pubblicazione: (2024)
When Wording Steers the Evaluation: Framing Bias in LLM judges
di: Hwang, Yerin, et al.
Pubblicazione: (2026)
di: Hwang, Yerin, et al.
Pubblicazione: (2026)
KITE: A Benchmark for Evaluating Korean Instruction-Following Abilities in Large Language Models
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
di: Kim, Dongjun, et al.
Pubblicazione: (2025)
SAIF: A Comprehensive Framework for Evaluating the Risks of Generative AI in the Public Sector
di: Lee, Kyeongryul, et al.
Pubblicazione: (2025)
di: Lee, Kyeongryul, et al.
Pubblicazione: (2025)
TextAtari: 100K Frames Game Playing with Language Agents
di: Li, Wenhao, et al.
Pubblicazione: (2025)
di: Li, Wenhao, et al.
Pubblicazione: (2025)
Debiasing Large Language Models toward Social Factors in Online Behavior Analytics through Prompt Knowledge Tuning
di: Salemi, Hossein, et al.
Pubblicazione: (2026)
di: Salemi, Hossein, et al.
Pubblicazione: (2026)
Large, Small or Both: A Novel Data Augmentation Framework Based on Language Models for Debiasing Opinion Summarization
di: Zhang, Yanyue, et al.
Pubblicazione: (2024)
di: Zhang, Yanyue, et al.
Pubblicazione: (2024)
Towards Transparency: Exploring LLM Trainings Datasets through Visual Topic Modeling and Semantic Frame
di: de Dampierre, Charles, et al.
Pubblicazione: (2024)
di: de Dampierre, Charles, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ERBench: An Entity-Relationship based Automatically Verifiable Hallucination Benchmark for Large Language Models
di: Oh, Jio, et al.
Pubblicazione: (2024) -
Classroom AI: Large Language Models as Grade-Specific Teachers
di: Oh, Jio, et al.
Pubblicazione: (2026) -
Harnessing Temporal Databases for Systematic Evaluation of Factual Time-Sensitive Question-Answering in Large Language Models
di: Kim, Soyeon, et al.
Pubblicazione: (2025) -
PFGuard: A Generative Framework with Privacy and Fairness Safeguards
di: Kim, Soyeon, et al.
Pubblicazione: (2024) -
DialectLLM: A Dialect-Aware Dialog[ue] Generation Framework Beyond Standard American English
di: Oh, Jio, et al.
Pubblicazione: (2026)