Causal-Guided Active Learning for Debiasing Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Du, Li, Sun, Zhouhao, Ding, Xiao, Ma, Yixuan, Zhao, Yang, Qiu, Kaitao, Liu, Ting, Qin, Bing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Information Gain-Guided Causal Intervention for Autonomous Debiasing Large Language Models
di: Sun, Zhouhao, et al.
Pubblicazione: (2025)
di: Sun, Zhouhao, et al.
Pubblicazione: (2025)
Large Language Models Are Still Misled by Simple Bias Ensembles
di: Sun, Zhouhao, et al.
Pubblicazione: (2025)
di: Sun, Zhouhao, et al.
Pubblicazione: (2025)
Deciphering the Impact of Pretraining Data on Large Language Models through Machine Unlearning
di: Zhao, Yang, et al.
Pubblicazione: (2024)
di: Zhao, Yang, et al.
Pubblicazione: (2024)
Towards Generalizable and Faithful Logic Reasoning over Natural Language via Resolution Refutation
di: Sun, Zhouhao, et al.
Pubblicazione: (2024)
di: Sun, Zhouhao, et al.
Pubblicazione: (2024)
Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models
di: Xiong, Kai, et al.
Pubblicazione: (2025)
di: Xiong, Kai, et al.
Pubblicazione: (2025)
Self-Route: Automatic Mode Switching via Capability Estimation for Efficient Reasoning
di: He, Yang, et al.
Pubblicazione: (2025)
di: He, Yang, et al.
Pubblicazione: (2025)
Examining Inter-Consistency of Large Language Models Collaboration: An In-depth Analysis via Debate
di: Xiong, Kai, et al.
Pubblicazione: (2023)
di: Xiong, Kai, et al.
Pubblicazione: (2023)
LLM4Rec: Large Language Models for Multimodal Generative Recommendation with Causal Debiasing
di: Ma, Bo, et al.
Pubblicazione: (2025)
di: Ma, Bo, et al.
Pubblicazione: (2025)
Meaningful Learning: Enhancing Abstract Reasoning in Large Language Models via Generic Fact Guidance
di: Xiong, Kai, et al.
Pubblicazione: (2024)
di: Xiong, Kai, et al.
Pubblicazione: (2024)
GR-Ben: A General Reasoning Benchmark for Evaluating Process Reward Models
di: Sun, Zhouhao, et al.
Pubblicazione: (2026)
di: Sun, Zhouhao, et al.
Pubblicazione: (2026)
Debiasing Large Language Models via Adaptive Causal Prompting with Sketch-of-Thought
di: Li, Bowen, et al.
Pubblicazione: (2026)
di: Li, Bowen, et al.
Pubblicazione: (2026)
Bi-directional Bias Attribution: Debiasing Large Language Models without Modifying Prompts
di: Lin, Yujie, et al.
Pubblicazione: (2026)
di: Lin, Yujie, et al.
Pubblicazione: (2026)
UGID: Unified Graph Isomorphism for Debiasing Large Language Models
di: Ding, Zikang, et al.
Pubblicazione: (2026)
di: Ding, Zikang, et al.
Pubblicazione: (2026)
Enhancing Complex Causality Extraction via Improved Subtask Interaction and Knowledge Fusion
di: Gao, Jinglong, et al.
Pubblicazione: (2024)
di: Gao, Jinglong, et al.
Pubblicazione: (2024)
Continual Learning Using Only Large Language Model Prompting
di: Qiu, Jiabao, et al.
Pubblicazione: (2024)
di: Qiu, Jiabao, et al.
Pubblicazione: (2024)
ReCo: Reliable Causal Chain Reasoning via Structural Causal Recurrent Neural Networks
di: Xiong, Kai, et al.
Pubblicazione: (2022)
di: Xiong, Kai, et al.
Pubblicazione: (2022)
Diagnosing and Remedying Knowledge Deficiencies in LLMs via Label-free Curricular Meaningful Learning
di: Xiong, Kai, et al.
Pubblicazione: (2024)
di: Xiong, Kai, et al.
Pubblicazione: (2024)
ELAD: Explanation-Guided Large Language Models Active Distillation
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
di: Zhang, Yifei, et al.
Pubblicazione: (2024)
Safety-Aware Fine-Tuning of Large Language Models
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
DeepTool: Scaling Interleaved Deliberation in Tool-Integrated Reasoning via Process-Supervised Reinforcement Learning
di: He, Yang, et al.
Pubblicazione: (2026)
di: He, Yang, et al.
Pubblicazione: (2026)
Causal Inference with Large Language Model: A Survey
di: Ma, Jing
Pubblicazione: (2024)
di: Ma, Jing
Pubblicazione: (2024)
Advancing Large Language Model Attribution through Self-Improving
di: Huang, Lei, et al.
Pubblicazione: (2024)
di: Huang, Lei, et al.
Pubblicazione: (2024)
AdaThink-Med: Medical Adaptive Thinking with Uncertainty-Guided Length Calibration
di: Rui, Shaohao, et al.
Pubblicazione: (2025)
di: Rui, Shaohao, et al.
Pubblicazione: (2025)
Active Use of Latent Constituency Representation in both Humans and Large Language Models
di: Liu, Wei, et al.
Pubblicazione: (2024)
di: Liu, Wei, et al.
Pubblicazione: (2024)
Ontology-Guided Reverse Thinking Makes Large Language Models Stronger on Knowledge Graph Question Answering
di: Liu, Runxuan, et al.
Pubblicazione: (2025)
di: Liu, Runxuan, et al.
Pubblicazione: (2025)
General Phrase Debiaser: Debiasing Masked Language Models at a Multi-Token Level
di: Shi, Bingkang, et al.
Pubblicazione: (2023)
di: Shi, Bingkang, et al.
Pubblicazione: (2023)
UFO-RL: Uncertainty-Focused Optimization for Efficient Reinforcement Learning Data Selection
di: Zhao, Yang, et al.
Pubblicazione: (2025)
di: Zhao, Yang, et al.
Pubblicazione: (2025)
Take its Essence, Discard its Dross! Debiasing for Toxic Language Detection via Counterfactual Causal Effect
di: Lu, Junyu, et al.
Pubblicazione: (2024)
di: Lu, Junyu, et al.
Pubblicazione: (2024)
DeFrame: Debiasing Large Language Models Against Framing Effects
di: Lim, Kahee, et al.
Pubblicazione: (2026)
di: Lim, Kahee, et al.
Pubblicazione: (2026)
Give Us the Facts: Enhancing Large Language Models with Knowledge Graphs for Fact-aware Language Modeling
di: Yang, Linyao, et al.
Pubblicazione: (2023)
di: Yang, Linyao, et al.
Pubblicazione: (2023)
MAPLE: A Framework for Active Preference Learning Guided by Large Language Models
di: Mahmud, Saaduddin, et al.
Pubblicazione: (2024)
di: Mahmud, Saaduddin, et al.
Pubblicazione: (2024)
BIPEFT: Budget-Guided Iterative Search for Parameter Efficient Fine-Tuning of Large Pretrained Language Models
di: Chang, Aofei, et al.
Pubblicazione: (2024)
di: Chang, Aofei, et al.
Pubblicazione: (2024)
Self-Supervised Position Debiasing for Large Language Models
di: Liu, Zhongkun, et al.
Pubblicazione: (2024)
di: Liu, Zhongkun, et al.
Pubblicazione: (2024)
From AR to Diffusion: Efficiently Adapting Large Language Models with Strictly Causal and Elastic Horizons
di: Ma, Xiangyu, et al.
Pubblicazione: (2026)
di: Ma, Xiangyu, et al.
Pubblicazione: (2026)
Debiasing Reward Models via Causally Motivated Inference-Time Intervention
di: Shinoda, Kazutoshi, et al.
Pubblicazione: (2026)
di: Shinoda, Kazutoshi, et al.
Pubblicazione: (2026)
GraCoRe: Benchmarking Graph Comprehension and Complex Reasoning in Large Language Models
di: Yuan, Zike, et al.
Pubblicazione: (2024)
di: Yuan, Zike, et al.
Pubblicazione: (2024)
Causal Agent based on Large Language Model
di: Han, Kairong, et al.
Pubblicazione: (2024)
di: Han, Kairong, et al.
Pubblicazione: (2024)
Self-Evolving GPT: A Lifelong Autonomous Experiential Learner
di: Gao, Jinglong, et al.
Pubblicazione: (2024)
di: Gao, Jinglong, et al.
Pubblicazione: (2024)
Knowledge-tuning Large Language Models with Structured Medical Knowledge Bases for Reliable Response Generation in Chinese
di: Wang, Haochun, et al.
Pubblicazione: (2023)
di: Wang, Haochun, et al.
Pubblicazione: (2023)
PICLe: Eliciting Diverse Behaviors from Large Language Models with Persona In-Context Learning
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
di: Choi, Hyeong Kyu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Information Gain-Guided Causal Intervention for Autonomous Debiasing Large Language Models
di: Sun, Zhouhao, et al.
Pubblicazione: (2025) -
Large Language Models Are Still Misled by Simple Bias Ensembles
di: Sun, Zhouhao, et al.
Pubblicazione: (2025) -
Deciphering the Impact of Pretraining Data on Large Language Models through Machine Unlearning
di: Zhao, Yang, et al.
Pubblicazione: (2024) -
Towards Generalizable and Faithful Logic Reasoning over Natural Language via Resolution Refutation
di: Sun, Zhouhao, et al.
Pubblicazione: (2024) -
Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models
di: Xiong, Kai, et al.
Pubblicazione: (2025)