Q-function Decomposition with Intervention Semantics with Factored Action Spaces
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Junkyu, Gao, Tian, Nelson, Elliot, Liu, Miao, Bhattacharjya, Debarun, Lu, Songtao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Consistency Hypothesis in Uncertainty Quantification for Large Language Models
by: Xiao, Quan, et al.
Published: (2025)
by: Xiao, Quan, et al.
Published: (2025)
Foundation Model Sherpas: Guiding Foundation Models through Knowledge and Reasoning
by: Bhattacharjya, Debarun, et al.
Published: (2024)
by: Bhattacharjya, Debarun, et al.
Published: (2024)
Agentics 2.0: Logical Transduction Algebra for Agentic Data Workflows
by: Gliozzo, Alfio Massimiliano, et al.
Published: (2026)
by: Gliozzo, Alfio Massimiliano, et al.
Published: (2026)
Knowledge Base Construction for Knowledge-Augmented Text-to-SQL
by: Baek, Jinheon, et al.
Published: (2025)
by: Baek, Jinheon, et al.
Published: (2025)
SIMBA UQ: Similarity-Based Aggregation for Uncertainty Quantification in Large Language Models
by: Bhattacharjya, Debarun, et al.
Published: (2025)
by: Bhattacharjya, Debarun, et al.
Published: (2025)
Model-based Reinforcement Learning for Parameterized Action Spaces
by: Zhang, Renhao, et al.
Published: (2024)
by: Zhang, Renhao, et al.
Published: (2024)
TimeKAN: KAN-based Frequency Decomposition Learning Architecture for Long-term Time Series Forecasting
by: Huang, Songtao, et al.
Published: (2025)
by: Huang, Songtao, et al.
Published: (2025)
Stochastic Q-learning for Large Discrete Action Spaces
by: Fourati, Fares, et al.
Published: (2024)
by: Fourati, Fares, et al.
Published: (2024)
Divergence of Empirical Neural Tangent Kernel in Classification Problems
by: Yu, Zixiong, et al.
Published: (2025)
by: Yu, Zixiong, et al.
Published: (2025)
FactReasoner: A Probabilistic Approach to Long-Form Factuality Assessment for Large Language Models
by: Marinescu, Radu, et al.
Published: (2025)
by: Marinescu, Radu, et al.
Published: (2025)
A Single-Loop Gradient Descent and Perturbed Ascent Algorithm for Nonconvex Functional Constrained Optimization
by: Lu, Songtao
Published: (2022)
by: Lu, Songtao
Published: (2022)
Transduction is All You Need for Structured Data Workflows
by: Gliozzo, Alfio, et al.
Published: (2025)
by: Gliozzo, Alfio, et al.
Published: (2025)
From Isolated Islands to Pangea: Unifying Semantic Space for Human Action Understanding
by: Li, Yong-Lu, et al.
Published: (2023)
by: Li, Yong-Lu, et al.
Published: (2023)
Self-Supervised Contrastive Pre-Training for Multivariate Point Processes
by: Shou, Xiao, et al.
Published: (2024)
by: Shou, Xiao, et al.
Published: (2024)
Q-LIME $π$: A Quantum-Inspired Extension to LIME
by: Vargas, Nelson Colón
Published: (2024)
by: Vargas, Nelson Colón
Published: (2024)
Reinforcing Language Agents via Policy Optimization with Action Decomposition
by: Wen, Muning, et al.
Published: (2024)
by: Wen, Muning, et al.
Published: (2024)
Action-Adaptive Continual Learning: Enabling Policy Generalization under Dynamic Action Spaces
by: Pan, Chaofan, et al.
Published: (2025)
by: Pan, Chaofan, et al.
Published: (2025)
Distributions as Actions: A Unified Framework for Diverse Action Spaces
by: He, Jiamin, et al.
Published: (2025)
by: He, Jiamin, et al.
Published: (2025)
Disentanglement in Difference: Directly Learning Semantically Disentangled Representations by Maximizing Inter-Factor Differences
by: Zhang, Xingshen, et al.
Published: (2025)
by: Zhang, Xingshen, et al.
Published: (2025)
Adaptive Action Chunking via Multi-Chunk Q Value Estimation
by: Shin, Yongjae, et al.
Published: (2026)
by: Shin, Yongjae, et al.
Published: (2026)
Robust Deep Reinforcement Learning with Adaptive Adversarial Perturbations in Action Space
by: Liu, Qianmei, et al.
Published: (2024)
by: Liu, Qianmei, et al.
Published: (2024)
Bilevel Optimization over Saddle Points of Zero-Sum Markov Games
by: Zheng, Zihao, et al.
Published: (2026)
by: Zheng, Zihao, et al.
Published: (2026)
Efficient Public Health Intervention Planning Using Decomposition-Based Decision-Focused Learning
by: Shah, Sanket, et al.
Published: (2024)
by: Shah, Sanket, et al.
Published: (2024)
Causal Discovery in Action: Learning Chain-Reaction Mechanisms from Interventions
by: Panayiotou, Panayiotis, et al.
Published: (2026)
by: Panayiotou, Panayiotis, et al.
Published: (2026)
HDNet: Physics-Inspired Neural Network for Flow Estimation based on Helmholtz Decomposition
by: Qi, Miao, et al.
Published: (2024)
by: Qi, Miao, et al.
Published: (2024)
Stochastic Parameter Decomposition
by: Bushnaq, Lucius, et al.
Published: (2025)
by: Bushnaq, Lucius, et al.
Published: (2025)
Time After Time: Deep-Q Effect Estimation for Interventions on When and What to do
by: Wald, Yoav, et al.
Published: (2025)
by: Wald, Yoav, et al.
Published: (2025)
Frictional Q-Learning
by: Kim, Hyunwoo, et al.
Published: (2025)
by: Kim, Hyunwoo, et al.
Published: (2025)
Local Reinforcement Learning with Action-Conditioned Root Mean Squared Q-Functions
by: Wu, Frank, et al.
Published: (2025)
by: Wu, Frank, et al.
Published: (2025)
A Framework for Quantifying How Pre-Training and Context Benefit In-Context Learning
by: Song, Bingqing, et al.
Published: (2025)
by: Song, Bingqing, et al.
Published: (2025)
Learning Complex Teamwork Tasks Using a Given Sub-task Decomposition
by: Fosong, Elliot, et al.
Published: (2023)
by: Fosong, Elliot, et al.
Published: (2023)
In-Context Reinforcement Learning for Variable Action Spaces
by: Sinii, Viacheslav, et al.
Published: (2023)
by: Sinii, Viacheslav, et al.
Published: (2023)
Low-Rank MDPs with Continuous Action Spaces
by: Bennett, Andrew, et al.
Published: (2023)
by: Bennett, Andrew, et al.
Published: (2023)
Balancing Multimodal Learning through Label Space Reshaping
by: Ma, Xiaoyu, et al.
Published: (2026)
by: Ma, Xiaoyu, et al.
Published: (2026)
Suppressing Overestimation in Q-Learning through Adversarial Behaviors
by: Lee, HyeAnn, et al.
Published: (2023)
by: Lee, HyeAnn, et al.
Published: (2023)
PIAD-SRNN: Physics-Informed Adaptive Decomposition in State-Space RNN
by: Mohammadshirazi, Ahmad, et al.
Published: (2024)
by: Mohammadshirazi, Ahmad, et al.
Published: (2024)
Chunk-Guided Q-Learning
by: Song, Gwanwoo, et al.
Published: (2026)
by: Song, Gwanwoo, et al.
Published: (2026)
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces
by: Tiwari, Saket, et al.
Published: (2022)
by: Tiwari, Saket, et al.
Published: (2022)
Periodic Regularized Q-Learning
by: Yang, Hyukjun, et al.
Published: (2026)
by: Yang, Hyukjun, et al.
Published: (2026)
Beyond Hidden-Layer Manipulation: Semantically-Aware Logit Interventions for Debiasing LLMs
by: Xia, Wei
Published: (2025)
by: Xia, Wei
Published: (2025)
Similar Items
-
The Consistency Hypothesis in Uncertainty Quantification for Large Language Models
by: Xiao, Quan, et al.
Published: (2025) -
Foundation Model Sherpas: Guiding Foundation Models through Knowledge and Reasoning
by: Bhattacharjya, Debarun, et al.
Published: (2024) -
Agentics 2.0: Logical Transduction Algebra for Agentic Data Workflows
by: Gliozzo, Alfio Massimiliano, et al.
Published: (2026) -
Knowledge Base Construction for Knowledge-Augmented Text-to-SQL
by: Baek, Jinheon, et al.
Published: (2025) -
SIMBA UQ: Similarity-Based Aggregation for Uncertainty Quantification in Large Language Models
by: Bhattacharjya, Debarun, et al.
Published: (2025)