Navigating Noisy Feedback: Enhancing Reinforcement Learning with Error-Prone Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, Muhan, Shi, Shuyang, Guo, Yue, Chalaki, Behdad, Tadiparthi, Vaishnav, Pari, Ehsan Moradi, Stepputtis, Simon, Campbell, Joseph, Sycara, Katia |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning
by: Lin, Muhan, et al.
Published: (2025)
by: Lin, Muhan, et al.
Published: (2025)
R3DM: Enabling Role Discovery and Diversity Through Dynamics Models in Multi-agent Reinforcement Learning
by: Goel, Harsh, et al.
Published: (2025)
by: Goel, Harsh, et al.
Published: (2025)
Language Grounded Multi-agent Reinforcement Learning with Human-interpretable Communication
by: Li, Huao, et al.
Published: (2024)
by: Li, Huao, et al.
Published: (2024)
In Search of a Lost Metric: Human Empowerment as a Pillar of Socially Conscious Navigation
by: Baddam, Vasanth Reddy, et al.
Published: (2025)
by: Baddam, Vasanth Reddy, et al.
Published: (2025)
SSR: A Generic Framework for Text-Aided Map Compression for Localization
by: Omama, Mohammad, et al.
Published: (2026)
by: Omama, Mohammad, et al.
Published: (2026)
Multi-Agent Transfer Learning via Temporal Contrastive Learning
by: Zeng, Weihao, et al.
Published: (2024)
by: Zeng, Weihao, et al.
Published: (2024)
HiKER-SGG: Hierarchical Knowledge Enhanced Robust Scene Graph Generation
by: Zhang, Ce, et al.
Published: (2024)
by: Zhang, Ce, et al.
Published: (2024)
SMART-Merge Planner: A Safe Merging and Real-Time Motion Planner for Autonomous Highway On-Ramp Merging
by: Mohammadnejad, Toktam, et al.
Published: (2025)
by: Mohammadnejad, Toktam, et al.
Published: (2025)
Dual Control for Interactive Autonomous Merging with Model Predictive Diffusion
by: Knaup, Jacob, et al.
Published: (2025)
by: Knaup, Jacob, et al.
Published: (2025)
Enhancing Vision-Language Few-Shot Adaptation with Negative Learning
by: Zhang, Ce, et al.
Published: (2024)
by: Zhang, Ce, et al.
Published: (2024)
Model-Agnostic Policy Explanations with Large Language Models
by: Xi-Jia, Zhang, et al.
Published: (2025)
by: Xi-Jia, Zhang, et al.
Published: (2025)
Dual Prototype Evolving for Test-Time Generalization of Vision-Language Models
by: Zhang, Ce, et al.
Published: (2024)
by: Zhang, Ce, et al.
Published: (2024)
GL-NeRF: Gauss-Laguerre Quadrature Enables Training-Free NeRF Acceleration
by: Yong, Silong, et al.
Published: (2024)
by: Yong, Silong, et al.
Published: (2024)
ShapeGrasp: Zero-Shot Task-Oriented Grasping with Large Language Models through Geometric Decomposition
by: Li, Samuel, et al.
Published: (2024)
by: Li, Samuel, et al.
Published: (2024)
Theory of Mind for Multi-Agent Collaboration via Large Language Models
by: Li, Huao, et al.
Published: (2023)
by: Li, Huao, et al.
Published: (2023)
Symbolic Graph Inference for Compound Scene Understanding
by: Aryan, FNU, et al.
Published: (2024)
by: Aryan, FNU, et al.
Published: (2024)
Learning Robust Reasoning through Guided Adversarial Self-Play
by: Li, Shuozhe, et al.
Published: (2026)
by: Li, Shuozhe, et al.
Published: (2026)
Spectral-Aware Global Fusion for RGB-Thermal Semantic Segmentation
by: Zhang, Ce, et al.
Published: (2025)
by: Zhang, Ce, et al.
Published: (2025)
SCALAR: Learning and Composing Skills through LLM Guided Symbolic Planning and Deep RL Grounding
by: Zabounidis, Renos, et al.
Published: (2026)
by: Zabounidis, Renos, et al.
Published: (2026)
Modeling Latent Partner Strategies for Adaptive Zero-Shot Human-Agent Collaboration
by: Li, Benjamin, et al.
Published: (2025)
by: Li, Benjamin, et al.
Published: (2025)
Adaptively Coordinating with Novel Partners via Learned Latent Strategies
by: Li, Benjamin, et al.
Published: (2025)
by: Li, Benjamin, et al.
Published: (2025)
Metacognitive Self-Correction for Multi-Agent System via Prototype-Guided Next-Execution Reconstruction
by: Shen, Xu, et al.
Published: (2025)
by: Shen, Xu, et al.
Published: (2025)
B3C: A Minimalist Approach to Offline Multi-Agent Reinforcement Learning
by: Kim, Woojun, et al.
Published: (2025)
by: Kim, Woojun, et al.
Published: (2025)
Overcoming Valid Action Suppression in Unmasked Policy Gradient Algorithms
by: Zabounidis, Renos, et al.
Published: (2026)
by: Zabounidis, Renos, et al.
Published: (2026)
Theory of Mind Guided Strategy Adaptation for Zero-Shot Coordination
by: Ni, Andrew, et al.
Published: (2026)
by: Ni, Andrew, et al.
Published: (2026)
Energy-Based Transfer for Reinforcement Learning
by: Deng, Zeyun, et al.
Published: (2025)
by: Deng, Zeyun, et al.
Published: (2025)
Sigma: Siamese Mamba Network for Multi-Modal Semantic Segmentation
by: Wan, Zifu, et al.
Published: (2024)
by: Wan, Zifu, et al.
Published: (2024)
HiMemFormer: Hierarchical Memory-Aware Transformer for Multi-Agent Action Anticipation
by: Wang, Zirui, et al.
Published: (2024)
by: Wang, Zirui, et al.
Published: (2024)
Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models
by: Zhang, Ce, et al.
Published: (2025)
by: Zhang, Ce, et al.
Published: (2025)
CARE: Enhancing Safety of Visual Navigation through Collision Avoidance via Repulsive Estimation
by: Kim, Joonkyung, et al.
Published: (2025)
by: Kim, Joonkyung, et al.
Published: (2025)
OMG: Opacity Matters in Material Modeling with Gaussian Splatting
by: Yong, Silong, et al.
Published: (2025)
by: Yong, Silong, et al.
Published: (2025)
InstructPart: Task-Oriented Part Segmentation with Instruction Reasoning
by: Wan, Zifu, et al.
Published: (2025)
by: Wan, Zifu, et al.
Published: (2025)
Multi-Robot Navigation in Social Mini-Games: Definitions, Taxonomy, and Algorithms
by: Chandra, Rohan, et al.
Published: (2025)
by: Chandra, Rohan, et al.
Published: (2025)
Stochastic Time-Optimal Trajectory Planning for Connected and Automated Vehicles in Mixed-Traffic Merging Scenarios
by: Le, Viet-Anh, et al.
Published: (2023)
by: Le, Viet-Anh, et al.
Published: (2023)
Fair Cooperation in Mixed-Motive Games via Conflict-Aware Gradient Adjustment
by: Kim, Woojun, et al.
Published: (2025)
by: Kim, Woojun, et al.
Published: (2025)
AROW: V2X-based Automated Right-of-Way Algorithm for Cooperative Intersection Management
by: Shah, Ghayoor, et al.
Published: (2023)
by: Shah, Ghayoor, et al.
Published: (2023)
BET: Explaining Deep Reinforcement Learning through The Error-Prone Decisions
by: Liu, Xiao, et al.
Published: (2024)
by: Liu, Xiao, et al.
Published: (2024)
ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models
by: Wan, Zifu, et al.
Published: (2025)
by: Wan, Zifu, et al.
Published: (2025)
SubTokenTest: A Practical Benchmark for Real-World Sub-token Understanding
by: Hou, Shuyang, et al.
Published: (2026)
by: Hou, Shuyang, et al.
Published: (2026)
A Research and Educational Robotic Testbed for Real-time Control of Emerging Mobility Systems: From Theory to Scaled Experiments
by: Chalaki, Behdad, et al.
Published: (2021)
by: Chalaki, Behdad, et al.
Published: (2021)
Similar Items
-
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning
by: Lin, Muhan, et al.
Published: (2025) -
R3DM: Enabling Role Discovery and Diversity Through Dynamics Models in Multi-agent Reinforcement Learning
by: Goel, Harsh, et al.
Published: (2025) -
Language Grounded Multi-agent Reinforcement Learning with Human-interpretable Communication
by: Li, Huao, et al.
Published: (2024) -
In Search of a Lost Metric: Human Empowerment as a Pillar of Socially Conscious Navigation
by: Baddam, Vasanth Reddy, et al.
Published: (2025) -
SSR: A Generic Framework for Text-Aided Map Compression for Localization
by: Omama, Mohammad, et al.
Published: (2026)