SCALAR: Learning and Composing Skills through LLM Guided Symbolic Planning and Deep RL Grounding
Fuente:
arXiv
Salvato in:
| Autori principali: | Zabounidis, Renos, Wu, Yue, Stepputtis, Simon, Kim, Woojun, Li, Yuanzhi, Mitchell, Tom, Sycara, Katia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Overcoming Valid Action Suppression in Unmasked Policy Gradient Algorithms
di: Zabounidis, Renos, et al.
Pubblicazione: (2026)
di: Zabounidis, Renos, et al.
Pubblicazione: (2026)
B3C: A Minimalist Approach to Offline Multi-Agent Reinforcement Learning
di: Kim, Woojun, et al.
Pubblicazione: (2025)
di: Kim, Woojun, et al.
Pubblicazione: (2025)
Theory of Mind Guided Strategy Adaptation for Zero-Shot Coordination
di: Ni, Andrew, et al.
Pubblicazione: (2026)
di: Ni, Andrew, et al.
Pubblicazione: (2026)
Fair Cooperation in Mixed-Motive Games via Conflict-Aware Gradient Adjustment
di: Kim, Woojun, et al.
Pubblicazione: (2025)
di: Kim, Woojun, et al.
Pubblicazione: (2025)
DyPNIPP: Predicting Environment Dynamics for RL-based Robust Informative Path Planning
di: Deolasee, Srujan, et al.
Pubblicazione: (2024)
di: Deolasee, Srujan, et al.
Pubblicazione: (2024)
ShapeGrasp: Zero-Shot Task-Oriented Grasping with Large Language Models through Geometric Decomposition
di: Li, Samuel, et al.
Pubblicazione: (2024)
di: Li, Samuel, et al.
Pubblicazione: (2024)
Speaking the Language of Teamwork: LLM-Guided Credit Assignment in Multi-Agent Reinforcement Learning
di: Lin, Muhan, et al.
Pubblicazione: (2025)
di: Lin, Muhan, et al.
Pubblicazione: (2025)
HiMemFormer: Hierarchical Memory-Aware Transformer for Multi-Agent Action Anticipation
di: Wang, Zirui, et al.
Pubblicazione: (2024)
di: Wang, Zirui, et al.
Pubblicazione: (2024)
OffRIPP: Offline RL-based Informative Path Planning
di: Gadipudi, Srikar Babu, et al.
Pubblicazione: (2024)
di: Gadipudi, Srikar Babu, et al.
Pubblicazione: (2024)
Multi-Agent Transfer Learning via Temporal Contrastive Learning
di: Zeng, Weihao, et al.
Pubblicazione: (2024)
di: Zeng, Weihao, et al.
Pubblicazione: (2024)
Enhancing Vision-Language Few-Shot Adaptation with Negative Learning
di: Zhang, Ce, et al.
Pubblicazione: (2024)
di: Zhang, Ce, et al.
Pubblicazione: (2024)
Aligning LLM+PDDL Symbolic Plans with Human Objective Specifications through Evolutionary Algorithm Guidance
di: Burns, Owen, et al.
Pubblicazione: (2024)
di: Burns, Owen, et al.
Pubblicazione: (2024)
CARE: Enhancing Safety of Visual Navigation through Collision Avoidance via Repulsive Estimation
di: Kim, Joonkyung, et al.
Pubblicazione: (2025)
di: Kim, Joonkyung, et al.
Pubblicazione: (2025)
Dual Prototype Evolving for Test-Time Generalization of Vision-Language Models
di: Zhang, Ce, et al.
Pubblicazione: (2024)
di: Zhang, Ce, et al.
Pubblicazione: (2024)
GL-NeRF: Gauss-Laguerre Quadrature Enables Training-Free NeRF Acceleration
di: Yong, Silong, et al.
Pubblicazione: (2024)
di: Yong, Silong, et al.
Pubblicazione: (2024)
Adaptively Coordinating with Novel Partners via Learned Latent Strategies
di: Li, Benjamin, et al.
Pubblicazione: (2025)
di: Li, Benjamin, et al.
Pubblicazione: (2025)
Symbolic Graph Inference for Compound Scene Understanding
di: Aryan, FNU, et al.
Pubblicazione: (2024)
di: Aryan, FNU, et al.
Pubblicazione: (2024)
Spectral-Aware Global Fusion for RGB-Thermal Semantic Segmentation
di: Zhang, Ce, et al.
Pubblicazione: (2025)
di: Zhang, Ce, et al.
Pubblicazione: (2025)
HiKER-SGG: Hierarchical Knowledge Enhanced Robust Scene Graph Generation
di: Zhang, Ce, et al.
Pubblicazione: (2024)
di: Zhang, Ce, et al.
Pubblicazione: (2024)
Modeling Latent Partner Strategies for Adaptive Zero-Shot Human-Agent Collaboration
di: Li, Benjamin, et al.
Pubblicazione: (2025)
di: Li, Benjamin, et al.
Pubblicazione: (2025)
CDE: Concept-Driven Exploration for Reinforcement Learning
di: Mao, Le, et al.
Pubblicazione: (2025)
di: Mao, Le, et al.
Pubblicazione: (2025)
Re-FORC: Adaptive Reward Prediction for Efficient Chain-of-Thought Reasoning
di: Zabounidis, Renos, et al.
Pubblicazione: (2025)
di: Zabounidis, Renos, et al.
Pubblicazione: (2025)
Distributed Multi-robot Source Seeking in Unknown Environments with Unknown Number of Sources
di: Chen, Lingpeng, et al.
Pubblicazione: (2025)
di: Chen, Lingpeng, et al.
Pubblicazione: (2025)
Model-Agnostic Policy Explanations with Large Language Models
di: Xi-Jia, Zhang, et al.
Pubblicazione: (2025)
di: Xi-Jia, Zhang, et al.
Pubblicazione: (2025)
Sigma: Siamese Mamba Network for Multi-Modal Semantic Segmentation
di: Wan, Zifu, et al.
Pubblicazione: (2024)
di: Wan, Zifu, et al.
Pubblicazione: (2024)
Navigating Noisy Feedback: Enhancing Reinforcement Learning with Error-Prone Language Models
di: Lin, Muhan, et al.
Pubblicazione: (2024)
di: Lin, Muhan, et al.
Pubblicazione: (2024)
Theory of Mind for Multi-Agent Collaboration via Large Language Models
di: Li, Huao, et al.
Pubblicazione: (2023)
di: Li, Huao, et al.
Pubblicazione: (2023)
From Local Corrections to Generalized Skills: Improving Neuro-Symbolic Policies with MEMO
di: Christie, Benjamin A., et al.
Pubblicazione: (2026)
di: Christie, Benjamin A., et al.
Pubblicazione: (2026)
OMG: Opacity Matters in Material Modeling with Gaussian Splatting
di: Yong, Silong, et al.
Pubblicazione: (2025)
di: Yong, Silong, et al.
Pubblicazione: (2025)
InstructPart: Task-Oriented Part Segmentation with Instruction Reasoning
di: Wan, Zifu, et al.
Pubblicazione: (2025)
di: Wan, Zifu, et al.
Pubblicazione: (2025)
Optimal Task Assignment and Path Planning using Conflict-Based Search with Precedence and Temporal Constraints
di: Chong, Yu Quan, et al.
Pubblicazione: (2024)
di: Chong, Yu Quan, et al.
Pubblicazione: (2024)
LogiCity: Advancing Neuro-Symbolic AI with Abstract Urban Simulation
di: Li, Bowen, et al.
Pubblicazione: (2024)
di: Li, Bowen, et al.
Pubblicazione: (2024)
SmartPlay: A Benchmark for LLMs as Intelligent Agents
di: Wu, Yue, et al.
Pubblicazione: (2023)
di: Wu, Yue, et al.
Pubblicazione: (2023)
SCALAR: Scale-wise Controllable Visual Autoregressive Learning
di: Xu, Ryan, et al.
Pubblicazione: (2025)
di: Xu, Ryan, et al.
Pubblicazione: (2025)
LESSON: Learning to Integrate Exploration Strategies for Reinforcement Learning via an Option Framework
di: Kim, Woojun, et al.
Pubblicazione: (2023)
di: Kim, Woojun, et al.
Pubblicazione: (2023)
ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models
di: Wan, Zifu, et al.
Pubblicazione: (2025)
di: Wan, Zifu, et al.
Pubblicazione: (2025)
Reconfigurable Robot Control Using Flexible Coupling Mechanisms
di: Yi, Sha, et al.
Pubblicazione: (2023)
di: Yi, Sha, et al.
Pubblicazione: (2023)
Jailbreaking Frontier Foundation Models Through Intention Deception
di: Wang, Xinhe, et al.
Pubblicazione: (2026)
di: Wang, Xinhe, et al.
Pubblicazione: (2026)
CBGT-Net: A Neuromimetic Architecture for Robust Classification of Streaming Data
di: Sharma, Shreya, et al.
Pubblicazione: (2024)
di: Sharma, Shreya, et al.
Pubblicazione: (2024)
SCALAR: Self-Calibrating Adaptive Latent Attention Representation Learning
di: Abbas, Farwa, et al.
Pubblicazione: (2025)
di: Abbas, Farwa, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Overcoming Valid Action Suppression in Unmasked Policy Gradient Algorithms
di: Zabounidis, Renos, et al.
Pubblicazione: (2026) -
B3C: A Minimalist Approach to Offline Multi-Agent Reinforcement Learning
di: Kim, Woojun, et al.
Pubblicazione: (2025) -
Theory of Mind Guided Strategy Adaptation for Zero-Shot Coordination
di: Ni, Andrew, et al.
Pubblicazione: (2026) -
Fair Cooperation in Mixed-Motive Games via Conflict-Aware Gradient Adjustment
di: Kim, Woojun, et al.
Pubblicazione: (2025) -
DyPNIPP: Predicting Environment Dynamics for RL-based Robust Informative Path Planning
di: Deolasee, Srujan, et al.
Pubblicazione: (2024)