Probing Implicit Bias in Semi-gradient Q-learning: Visualizing the Effective Loss Landscapes via the Fokker--Planck Equation
Fuente:
arXiv
Salvato in:
| Autori principali: | Yin, Shuyu, Wen, Fei, Liu, Peilin, Luo, Tao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A priori Estimates for Deep Residual Network in Continuous-time Reinforcement Learning
di: Yin, Shuyu, et al.
Pubblicazione: (2024)
di: Yin, Shuyu, et al.
Pubblicazione: (2024)
Semi-Implicit Neural Ordinary Differential Equations
di: Zhang, Hong, et al.
Pubblicazione: (2024)
di: Zhang, Hong, et al.
Pubblicazione: (2024)
Score-Based Physics-Informed Neural Networks for High-Dimensional Fokker-Planck Equations
di: Hu, Zheyuan, et al.
Pubblicazione: (2024)
di: Hu, Zheyuan, et al.
Pubblicazione: (2024)
Visualizing Loss Functions as Topological Landscape Profiles
di: Geniesse, Caleb, et al.
Pubblicazione: (2024)
di: Geniesse, Caleb, et al.
Pubblicazione: (2024)
Analyzing and Bridging the Gap between Maximizing Total Reward and Discounted Reward in Deep Reinforcement Learning
di: Yin, Shuyu, et al.
Pubblicazione: (2024)
di: Yin, Shuyu, et al.
Pubblicazione: (2024)
GenderBias-\emph{VL}: Benchmarking Gender Bias in Vision Language Models via Counterfactual Probing
di: Xiao, Yisong, et al.
Pubblicazione: (2024)
di: Xiao, Yisong, et al.
Pubblicazione: (2024)
FP-AbDiff: Improving Score-based Antibody Design by Capturing Nonequilibrium Dynamics through the Underlying Fokker-Planck Equation
di: Chen, Jiameng, et al.
Pubblicazione: (2025)
di: Chen, Jiameng, et al.
Pubblicazione: (2025)
QuantFPFlow: Quantum Amplitude Estimation for Fokker--Planck Policy Optimisation in Continuous Reinforcement Learning
di: Weinberg, Abraham Itzhak
Pubblicazione: (2026)
di: Weinberg, Abraham Itzhak
Pubblicazione: (2026)
Adapting Critic Match Loss Landscape Visualization to Off-policy Reinforcement Learning
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
Towards Instance-wise Personalized Federated Learning via Semi-Implicit Bayesian Prompt Tuning
di: Ye, Tiandi, et al.
Pubblicazione: (2025)
di: Ye, Tiandi, et al.
Pubblicazione: (2025)
Quotient Geometry, Effective Curvature, and Implicit Bias in Simple Shallow Neural Networks
di: Dong, Hang-Cheng, et al.
Pubblicazione: (2026)
di: Dong, Hang-Cheng, et al.
Pubblicazione: (2026)
Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
A Loss Landscape Visualization Framework for Interpreting Reinforcement Learning: An ADHDP Case Study
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
di: Liu, Jingyi, et al.
Pubblicazione: (2026)
DIAR: Diffusion-model-guided Implicit Q-learning with Adaptive Revaluation
di: Park, Jaehyun, et al.
Pubblicazione: (2024)
di: Park, Jaehyun, et al.
Pubblicazione: (2024)
Error Bounds for Physics-Informed Neural Networks in Fokker-Planck PDEs
di: Kong, Chun-Wei, et al.
Pubblicazione: (2024)
di: Kong, Chun-Wei, et al.
Pubblicazione: (2024)
There is a Singularity in the Loss Landscape
di: Lowell, Mark
Pubblicazione: (2022)
di: Lowell, Mark
Pubblicazione: (2022)
APEX: Probing Neural Networks via Activation Perturbation
di: Ren, Tao, et al.
Pubblicazione: (2026)
di: Ren, Tao, et al.
Pubblicazione: (2026)
Early-Warning Signals of Grokking via Loss-Landscape Geometry
di: Xu, Yongzhong
Pubblicazione: (2026)
di: Xu, Yongzhong
Pubblicazione: (2026)
Making Models Unmergeable via Scaling-Sensitive Loss Landscape
di: Jang, Minwoo, et al.
Pubblicazione: (2026)
di: Jang, Minwoo, et al.
Pubblicazione: (2026)
Implicit Bias in LLMs for Transgender Populations
di: Hirsch, Micaela, et al.
Pubblicazione: (2026)
di: Hirsch, Micaela, et al.
Pubblicazione: (2026)
Implicit Bias in LLMs: A Survey
di: Lin, Xinru, et al.
Pubblicazione: (2025)
di: Lin, Xinru, et al.
Pubblicazione: (2025)
StructBreak: Structural Cognitive Overload-Induced Safety Failures in MLLMs
di: Luo, Yang, et al.
Pubblicazione: (2026)
di: Luo, Yang, et al.
Pubblicazione: (2026)
On the Implicit Bias of Adam
di: Cattaneo, Matias D., et al.
Pubblicazione: (2023)
di: Cattaneo, Matias D., et al.
Pubblicazione: (2023)
Landscaper: Understanding Loss Landscapes Through Multi-Dimensional Topological Analysis
di: Chen, Jiaqing, et al.
Pubblicazione: (2026)
di: Chen, Jiaqing, et al.
Pubblicazione: (2026)
Implicit Bias of the JKO Scheme
di: Halmos, Peter, et al.
Pubblicazione: (2025)
di: Halmos, Peter, et al.
Pubblicazione: (2025)
Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
di: Wen, Yuchen, et al.
Pubblicazione: (2024)
di: Wen, Yuchen, et al.
Pubblicazione: (2024)
Implicit Bias-Like Patterns in Reasoning Models
di: Lee, Messi H. J., et al.
Pubblicazione: (2025)
di: Lee, Messi H. J., et al.
Pubblicazione: (2025)
Implicit Visual Bias Mitigation by Posterior Estimate Sharpening of a Bayesian Neural Network
di: Stone, Rebecca S, et al.
Pubblicazione: (2023)
di: Stone, Rebecca S, et al.
Pubblicazione: (2023)
Probing the Safety Response Boundary of Large Language Models via Unsafe Decoding Path Generation
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
Training Implicit Generative Models via an Invariant Statistical Loss
di: de Frutos, José Manuel, et al.
Pubblicazione: (2024)
di: de Frutos, José Manuel, et al.
Pubblicazione: (2024)
Visual SLAMMOT Considering Multiple Motion Models
di: Tian, Peilin, et al.
Pubblicazione: (2024)
di: Tian, Peilin, et al.
Pubblicazione: (2024)
It's All Connected: A Journey Through Test-Time Memorization, Attentional Bias, Retention, and Online Optimization
di: Behrouz, Ali, et al.
Pubblicazione: (2025)
di: Behrouz, Ali, et al.
Pubblicazione: (2025)
Epistemic Uncertainty-Weighted Loss for Visual Bias Mitigation
di: Stone, Rebecca S, et al.
Pubblicazione: (2022)
di: Stone, Rebecca S, et al.
Pubblicazione: (2022)
Disentangling Granularity: An Implicit Inductive Bias in Factorized VAEs
di: Chen, Zihao, et al.
Pubblicazione: (2025)
di: Chen, Zihao, et al.
Pubblicazione: (2025)
ImplicitBBQ: Benchmarking Implicit Bias in Large Language Models through Characteristic Based Cues
di: Vedula, Bhaskara Hanuma, et al.
Pubblicazione: (2026)
di: Vedula, Bhaskara Hanuma, et al.
Pubblicazione: (2026)
FP-IRL: Fokker--Planck Inverse Reinforcement Learning -- A Physics-Constrained Approach to Markov Decision Processes
di: Huang, Chengyang, et al.
Pubblicazione: (2023)
di: Huang, Chengyang, et al.
Pubblicazione: (2023)
A High-Throughput Spiking Neural Network Processor Enabling Synaptic Delay Emulation
di: Chen, Faquan, et al.
Pubblicazione: (2025)
di: Chen, Faquan, et al.
Pubblicazione: (2025)
Provably Mitigating Overoptimization in RLHF: Your SFT Loss is Implicitly an Adversarial Regularizer
di: Liu, Zhihan, et al.
Pubblicazione: (2024)
di: Liu, Zhihan, et al.
Pubblicazione: (2024)
STaR: Towards Effective and Stable Table Reasoning via Slow-Thinking Large Language Models
di: Zhang, Huajian, et al.
Pubblicazione: (2025)
di: Zhang, Huajian, et al.
Pubblicazione: (2025)
Model Merging on Loss Landscape: A Geometry Perspective
di: Lu, Juanwu, et al.
Pubblicazione: (2026)
di: Lu, Juanwu, et al.
Pubblicazione: (2026)
Documenti analoghi
-
A priori Estimates for Deep Residual Network in Continuous-time Reinforcement Learning
di: Yin, Shuyu, et al.
Pubblicazione: (2024) -
Semi-Implicit Neural Ordinary Differential Equations
di: Zhang, Hong, et al.
Pubblicazione: (2024) -
Score-Based Physics-Informed Neural Networks for High-Dimensional Fokker-Planck Equations
di: Hu, Zheyuan, et al.
Pubblicazione: (2024) -
Visualizing Loss Functions as Topological Landscape Profiles
di: Geniesse, Caleb, et al.
Pubblicazione: (2024) -
Analyzing and Bridging the Gap between Maximizing Total Reward and Discounted Reward in Deep Reinforcement Learning
di: Yin, Shuyu, et al.
Pubblicazione: (2024)