Mitigating Estimation Bias with Representation Learning in TD Error-Driven Regularization
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Haohui, Chen, Zhiyong, Liu, Aoxiang, Fang, Wentuo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Double Actor-Critic with TD Error-Driven Regularization in Reinforcement Learning
by: Chen, Haohui, et al.
Published: (2024)
by: Chen, Haohui, et al.
Published: (2024)
Mildly Conservative Regularized Evaluation for Offline Reinforcement Learning
by: Chen, Haohui, et al.
Published: (2025)
by: Chen, Haohui, et al.
Published: (2025)
Multi-State TD Target for Model-Free Reinforcement Learning
by: Wang, Wuhao, et al.
Published: (2024)
by: Wang, Wuhao, et al.
Published: (2024)
Relative Counterfactual Contrastive Learning for Mitigating Pretrained Stance Bias in Stance Detection
by: Zhang, Jiarui, et al.
Published: (2024)
by: Zhang, Jiarui, et al.
Published: (2024)
Multiagent Reinforcement Learning with Neighbor Action Estimation
by: Luo, Zhenglong, et al.
Published: (2026)
by: Luo, Zhenglong, et al.
Published: (2026)
Mitigating Degree Bias in Graph Representation Learning with Learnable Structural Augmentation and Structural Self-Attention
by: Hoang, Van Thuy, et al.
Published: (2025)
by: Hoang, Van Thuy, et al.
Published: (2025)
Mitigating Prior Errors in Causal Structure Learning: A Resilient Approach via Bayesian Networks
by: Chen, Lyuzhou, et al.
Published: (2023)
by: Chen, Lyuzhou, et al.
Published: (2023)
Implicit Federated In-context Learning For Task-Specific LLM Fine-Tuning
by: Li, Dongcheng, et al.
Published: (2025)
by: Li, Dongcheng, et al.
Published: (2025)
Hybrid DQN-TD3 Reinforcement Learning for Autonomous Navigation in Dynamic Environments
by: He, Xiaoyi, et al.
Published: (2025)
by: He, Xiaoyi, et al.
Published: (2025)
Rethinking KL Regularization in RLHF: From Value Estimation to Gradient Optimization
by: Liu, Kezhao, et al.
Published: (2025)
by: Liu, Kezhao, et al.
Published: (2025)
Bounds on Representation-Induced Confounding Bias for Treatment Effect Estimation
by: Melnychuk, Valentyn, et al.
Published: (2023)
by: Melnychuk, Valentyn, et al.
Published: (2023)
Exploring Adaptive MCTS with TD Learning in miniXCOM
by: Saadat, Kimiya, et al.
Published: (2022)
by: Saadat, Kimiya, et al.
Published: (2022)
What Does Flow Matching Bring To TD Learning?
by: Agrawalla, Bhavya, et al.
Published: (2026)
by: Agrawalla, Bhavya, et al.
Published: (2026)
Mitigating Participation Imbalance Bias in Asynchronous Federated Learning
by: Chang, Xiangyu, et al.
Published: (2025)
by: Chang, Xiangyu, et al.
Published: (2025)
Mitigating the Structural Bias in Graph Adversarial Defenses
by: Fang, Junyuan, et al.
Published: (2025)
by: Fang, Junyuan, et al.
Published: (2025)
Addressing Bias Through Ensemble Learning and Regularized Fine-Tuning
by: Radwan, Ahmed, et al.
Published: (2024)
by: Radwan, Ahmed, et al.
Published: (2024)
ErrorEraser: Unlearning Data Bias for Improved Continual Learning
by: Cao, Xuemei, et al.
Published: (2025)
by: Cao, Xuemei, et al.
Published: (2025)
Scalable Multiagent Reinforcement Learning with Collective Influence Estimation
by: Luo, Zhenglong, et al.
Published: (2026)
by: Luo, Zhenglong, et al.
Published: (2026)
A Convolution and Attention Based Encoder for Reinforcement Learning under Partial Observability
by: Wang, Wuhao, et al.
Published: (2025)
by: Wang, Wuhao, et al.
Published: (2025)
Label Deconvolution for Node Representation Learning on Large-scale Attributed Graphs against Learning Bias
by: Shi, Zhihao, et al.
Published: (2023)
by: Shi, Zhihao, et al.
Published: (2023)
Mitigating Reward Over-Optimization in RLHF via Behavior-Supported Regularization
by: Dai, Juntao, et al.
Published: (2025)
by: Dai, Juntao, et al.
Published: (2025)
Mitigating Gender Bias in Depression Detection via Counterfactual Inference
by: Hu, Mingxuan, et al.
Published: (2025)
by: Hu, Mingxuan, et al.
Published: (2025)
Adaptive Regularization of Representation Rank as an Implicit Constraint of Bellman Equation
by: He, Qiang, et al.
Published: (2024)
by: He, Qiang, et al.
Published: (2024)
Bias Fitting to Mitigate Length Bias of Reward Model in RLHF
by: Zhao, Kangwen, et al.
Published: (2025)
by: Zhao, Kangwen, et al.
Published: (2025)
Subgroups Matter for Robust Bias Mitigation
by: Alloula, Anissa, et al.
Published: (2025)
by: Alloula, Anissa, et al.
Published: (2025)
Mitigating Degree Bias Adaptively with Hard-to-Learn Nodes in Graph Contrastive Learning
by: Hu, Jingyu, et al.
Published: (2025)
by: Hu, Jingyu, et al.
Published: (2025)
Mitigating Exposure Bias in Score-Based Generation of Molecular Conformations
by: Wang, Sijia, et al.
Published: (2024)
by: Wang, Sijia, et al.
Published: (2024)
Representation-Driven Reinforcement Learning
by: Nabati, Ofir, et al.
Published: (2023)
by: Nabati, Ofir, et al.
Published: (2023)
Identification and Mitigating Bias in Quantum Machine Learning
by: Swaminathan, Nandhini, et al.
Published: (2024)
by: Swaminathan, Nandhini, et al.
Published: (2024)
Noise-Resilient Unsupervised Graph Representation Learning via Multi-Hop Feature Quality Estimation
by: Li, Shiyuan, et al.
Published: (2024)
by: Li, Shiyuan, et al.
Published: (2024)
Analysis of Off-Policy $n$-Step TD-Learning with Linear Function Approximation
by: Lim, Han-Dong, et al.
Published: (2025)
by: Lim, Han-Dong, et al.
Published: (2025)
IQL-TD-MPC: Implicit Q-Learning for Hierarchical Model Predictive Control
by: Chitnis, Rohan, et al.
Published: (2023)
by: Chitnis, Rohan, et al.
Published: (2023)
Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning
by: Falzari, Massimiliano, et al.
Published: (2025)
by: Falzari, Massimiliano, et al.
Published: (2025)
Efficient Bias Mitigation Without Privileged Information
by: Zarlenga, Mateo Espinosa, et al.
Published: (2024)
by: Zarlenga, Mateo Espinosa, et al.
Published: (2024)
SafePowerGraph-HIL: Real-Time HIL Validation of Heterogeneous GNNs for Bridging Sim-to-Real Gap in Power Grids
by: Ma, Aoxiang, et al.
Published: (2025)
by: Ma, Aoxiang, et al.
Published: (2025)
SafePowerGraph: Safety-aware Evaluation of Graph Neural Networks for Transmission Power Grids
by: Ghamizi, Salah, et al.
Published: (2024)
by: Ghamizi, Salah, et al.
Published: (2024)
Iterative Refinement Neural Operators are Learned Fixed-Point Solvers: A Principled Approach to Spectral Bias Mitigation
by: Liu, Xiaotian, et al.
Published: (2026)
by: Liu, Xiaotian, et al.
Published: (2026)
Whither Bias Goes, I Will Go: An Integrative, Systematic Review of Algorithmic Bias Mitigation
by: Hickman, Louis, et al.
Published: (2024)
by: Hickman, Louis, et al.
Published: (2024)
Bias-Restrained Prefix Representation Finetuning for Mathematical Reasoning
by: Liang, Sirui, et al.
Published: (2025)
by: Liang, Sirui, et al.
Published: (2025)
Provably Mitigating Overoptimization in RLHF: Your SFT Loss is Implicitly an Adversarial Regularizer
by: Liu, Zhihan, et al.
Published: (2024)
by: Liu, Zhihan, et al.
Published: (2024)
Similar Items
-
Double Actor-Critic with TD Error-Driven Regularization in Reinforcement Learning
by: Chen, Haohui, et al.
Published: (2024) -
Mildly Conservative Regularized Evaluation for Offline Reinforcement Learning
by: Chen, Haohui, et al.
Published: (2025) -
Multi-State TD Target for Model-Free Reinforcement Learning
by: Wang, Wuhao, et al.
Published: (2024) -
Relative Counterfactual Contrastive Learning for Mitigating Pretrained Stance Bias in Stance Detection
by: Zhang, Jiarui, et al.
Published: (2024) -
Multiagent Reinforcement Learning with Neighbor Action Estimation
by: Luo, Zhenglong, et al.
Published: (2026)