A Multi-Fidelity Control Variate Approach for Policy Gradient Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Xinjie, Neary, Cyrus, Gupta, Kushagra, Suttle, Wesley A., Ellis, Christian, Topcu, Ufuk, Fridovich-Keil, David |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bayesian Inverse Games with High-Dimensional Multi-Modal Observations
by: Jain, Yash, et al.
Published: (2026)
by: Jain, Yash, et al.
Published: (2026)
Linear-Quadratic Gaussian Games with Distributed Sparse Estimation
by: Qiu, Tianyu, et al.
Published: (2026)
by: Qiu, Tianyu, et al.
Published: (2026)
Second-Order Algorithms for Finding Local Nash Equilibria in Zero-Sum Games
by: Gupta, Kushagra, et al.
Published: (2024)
by: Gupta, Kushagra, et al.
Published: (2024)
Inferring Foresightedness in Dynamic Noncooperative Games
by: Armstrong, Cade, et al.
Published: (2024)
by: Armstrong, Cade, et al.
Published: (2024)
Auto-Encoding Bayesian Inverse Games
by: Liu, Xinjie, et al.
Published: (2024)
by: Liu, Xinjie, et al.
Published: (2024)
Deceptive Exploration in Multi-armed Bandits
by: Vurankaya, I. Arda, et al.
Published: (2025)
by: Vurankaya, I. Arda, et al.
Published: (2025)
Interleaved Information Structures in Dynamic Games: A General Framework with Application to the Linear-Quadratic Case
by: K, Janani S, et al.
Published: (2026)
by: K, Janani S, et al.
Published: (2026)
More Information is Not Always Better: Connections between Zero-Sum Local Nash Equilibria in Feedback and Open-Loop Information Patterns
by: Gupta, Kushagra, et al.
Published: (2025)
by: Gupta, Kushagra, et al.
Published: (2025)
Neural Port-Hamiltonian Differential Algebraic Equations for Compositional Learning of Electrical Networks
by: Neary, Cyrus, et al.
Published: (2024)
by: Neary, Cyrus, et al.
Published: (2024)
Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and Perception
by: Yang, Yunhao, et al.
Published: (2023)
by: Yang, Yunhao, et al.
Published: (2023)
Efficiently Solving Mixed-Hierarchy Games with Quasi-Policy Approximations
by: Khan, Hamzah, et al.
Published: (2026)
by: Khan, Hamzah, et al.
Published: (2026)
Hierarchical Control for Head-to-Head Autonomous Racing
by: Thakkar, Rishabh Saumil, et al.
Published: (2022)
by: Thakkar, Rishabh Saumil, et al.
Published: (2022)
Monotonic Transformation Invariant Multi-task Learning
by: Murthy, Surya, et al.
Published: (2025)
by: Murthy, Surya, et al.
Published: (2025)
When Should Agents Coordinate in Differentiable Sequential Decision Problems?
by: Probine, Caleb, et al.
Published: (2026)
by: Probine, Caleb, et al.
Published: (2026)
Approximate Feedback Nash Equilibria with Sparse Inter-Agent Dependencies
by: Liu, Xinjie, et al.
Published: (2024)
by: Liu, Xinjie, et al.
Published: (2024)
Game-theoretic Occlusion-Aware Motion Planning: an Efficient Hybrid-Information Approach
by: Gupta, Kushagra, et al.
Published: (2023)
by: Gupta, Kushagra, et al.
Published: (2023)
Improving Pre-Trained Vision-Language-Action Policies with Model-Based Search
by: Neary, Cyrus, et al.
Published: (2025)
by: Neary, Cyrus, et al.
Published: (2025)
Cooperative Bargaining Games Without Utilities: Mediated Solutions from Direction Oracles
by: Gupta, Kushagra, et al.
Published: (2025)
by: Gupta, Kushagra, et al.
Published: (2025)
V-VLAPS: Value-Guided Planning for Vision-Language-Action Models
by: Ren, Ke, et al.
Published: (2026)
by: Ren, Ke, et al.
Published: (2026)
IG-MCTS: Human-in-the-Loop Cooperative Navigation under Incomplete Information
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
UNCAP: Uncertainty-Guided Neurosymbolic Planning Using Natural Language Communication for Cooperative Autonomous Vehicles
by: Bhatt, Neel P., et al.
Published: (2025)
by: Bhatt, Neel P., et al.
Published: (2025)
Chance-Constrained Correlated Equilibria for Robust Noncooperative Coordination
by: Im, Jaehan, et al.
Published: (2026)
by: Im, Jaehan, et al.
Published: (2026)
Noncooperative Virtual Queue Coordination via Uncertainty-Aware Correlated Equilibria
by: Im, Jaehan, et al.
Published: (2026)
by: Im, Jaehan, et al.
Published: (2026)
Scalable Coordination with Chance-Constrained Correlated Equilibria via Reduced-Rank Structure
by: Im, Jaehan, et al.
Published: (2026)
by: Im, Jaehan, et al.
Published: (2026)
A Flow Matching Algorithm for Many-Shot Adaptation to Unseen Distributions
by: Ingebrand, Tyler, et al.
Published: (2026)
by: Ingebrand, Tyler, et al.
Published: (2026)
A Player Selection Network for Scalable Game-Theoretic Prediction and Planning
by: Qiu, Tianyu, et al.
Published: (2025)
by: Qiu, Tianyu, et al.
Published: (2025)
Robust Multi-Agent Reinforcement Learning for Small UAS Separation Assurance under GPS Degradation and Spoofing
by: Zongo, Alex, et al.
Published: (2026)
by: Zongo, Alex, et al.
Published: (2026)
Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
by: Patel, Bhrij, et al.
Published: (2023)
by: Patel, Bhrij, et al.
Published: (2023)
When Should a Leader Act Suboptimally? The Role of Inferability in Repeated Stackelberg Games
by: Karabag, Mustafa O., et al.
Published: (2023)
by: Karabag, Mustafa O., et al.
Published: (2023)
Generalized Information Gathering Under Dynamics Uncertainty
by: Palafox, Fernando, et al.
Published: (2026)
by: Palafox, Fernando, et al.
Published: (2026)
Act Natural! Extending Naturalistic Projection to Multimodal Behavior Scenarios
by: Khan, Hamzah I., et al.
Published: (2025)
by: Khan, Hamzah I., et al.
Published: (2025)
Flow Policy Gradients for Robot Control
by: Yi, Brent, et al.
Published: (2026)
by: Yi, Brent, et al.
Published: (2026)
Zero to Autonomy in Real-Time: Online Adaptation of Dynamics in Unstructured Environments
by: Ward, William, et al.
Published: (2025)
by: Ward, William, et al.
Published: (2025)
Online Adaptation of Terrain-Aware Dynamics for Planning in Unstructured Environments
by: Ward, William, et al.
Published: (2025)
by: Ward, William, et al.
Published: (2025)
Coordination in Noncooperative Multiplayer Matrix Games via Reduced Rank Correlated Equilibria
by: Im, Jaehan, et al.
Published: (2024)
by: Im, Jaehan, et al.
Published: (2024)
Secure Coordination for Vertiport Sequencing in Advanced Air Mobility
by: Im, Jaehan, et al.
Published: (2026)
by: Im, Jaehan, et al.
Published: (2026)
Game-theoretic Decentralized Coordination for Airspace Sector Overload Mitigation
by: Im, Jaehan, et al.
Published: (2025)
by: Im, Jaehan, et al.
Published: (2025)
Simulation Distillation: Pretraining World Models in Simulation for Rapid Real-World Adaptation
by: Levy, Jacob, et al.
Published: (2026)
by: Levy, Jacob, et al.
Published: (2026)
Value of Information-based Deceptive Path Planning Under Adversarial Interventions
by: Suttle, Wesley A., et al.
Published: (2025)
by: Suttle, Wesley A., et al.
Published: (2025)
Online Foundation Model Selection in Robotics
by: Li, Po-han, et al.
Published: (2024)
by: Li, Po-han, et al.
Published: (2024)
Similar Items
-
Bayesian Inverse Games with High-Dimensional Multi-Modal Observations
by: Jain, Yash, et al.
Published: (2026) -
Linear-Quadratic Gaussian Games with Distributed Sparse Estimation
by: Qiu, Tianyu, et al.
Published: (2026) -
Second-Order Algorithms for Finding Local Nash Equilibria in Zero-Sum Games
by: Gupta, Kushagra, et al.
Published: (2024) -
Inferring Foresightedness in Dynamic Noncooperative Games
by: Armstrong, Cade, et al.
Published: (2024) -
Auto-Encoding Bayesian Inverse Games
by: Liu, Xinjie, et al.
Published: (2024)