CWM: Contrastive World Models for Action Feasibility Learning in Embodied Agent Pipelines
Fuente:
arXiv
Salvato in:
| Autore principale: | Banerjee, Chayan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Natural Language as Policies: Reasoning for Coordinate-Level Embodied Control with LLMs
di: Mikami, Yusuke, et al.
Pubblicazione: (2024)
di: Mikami, Yusuke, et al.
Pubblicazione: (2024)
BRAVE: Brain-Controlled Prosthetic Arm with Voice Integration and Embodied Learning for Enhanced Mobility
di: Basit, Abdul, et al.
Pubblicazione: (2025)
di: Basit, Abdul, et al.
Pubblicazione: (2025)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
LiloDriver: A Lifelong Learning Framework for Closed-loop Motion Planning in Long-tail Autonomous Driving Scenarios
di: Yao, Huaiyuan, et al.
Pubblicazione: (2025)
di: Yao, Huaiyuan, et al.
Pubblicazione: (2025)
MAP: Multi-user Personalization with Collaborative LLM-powered Agents
di: Lee, Christine, et al.
Pubblicazione: (2025)
di: Lee, Christine, et al.
Pubblicazione: (2025)
CSR: Infinite-Horizon Real-Time Policies with Massive Cached State Representations
di: Karlsson, Robin, et al.
Pubblicazione: (2026)
di: Karlsson, Robin, et al.
Pubblicazione: (2026)
EMOS: Embodiment-aware Heterogeneous Multi-robot Operating System with LLM Agents
di: Chen, Junting, et al.
Pubblicazione: (2024)
di: Chen, Junting, et al.
Pubblicazione: (2024)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
di: Li, Danyang, et al.
Pubblicazione: (2025)
di: Li, Danyang, et al.
Pubblicazione: (2025)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
Comparing Apples to Oranges: LLM-powered Multimodal Intention Prediction in an Object Categorization Task
di: Ali, Hassan, et al.
Pubblicazione: (2024)
di: Ali, Hassan, et al.
Pubblicazione: (2024)
Conversations with Andrea: Visitors' Opinions on Android Robots in a Museum
di: Heisler, Marcel, et al.
Pubblicazione: (2025)
di: Heisler, Marcel, et al.
Pubblicazione: (2025)
Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress
di: Agia, Christopher, et al.
Pubblicazione: (2024)
di: Agia, Christopher, et al.
Pubblicazione: (2024)
Tool-RoCo: An Agent-as-Tool Self-organization Large Language Model Benchmark in Multi-robot Cooperation
di: Zhang, Ke, et al.
Pubblicazione: (2025)
di: Zhang, Ke, et al.
Pubblicazione: (2025)
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
di: Luijkx, Jelle, et al.
Pubblicazione: (2025)
di: Luijkx, Jelle, et al.
Pubblicazione: (2025)
Speech-Gesture GAN: Gesture Generation for Robots and Embodied Agents
di: Liu, Carson Yu, et al.
Pubblicazione: (2023)
di: Liu, Carson Yu, et al.
Pubblicazione: (2023)
A Surveillance Based Interactive Robot
di: Kavimandan, Kshitij, et al.
Pubblicazione: (2025)
di: Kavimandan, Kshitij, et al.
Pubblicazione: (2025)
Autonomous Navigation and Collision Avoidance for Mobile Robots: Classification and Review
di: de Carvalho, Marcus Vinicius Leal, et al.
Pubblicazione: (2024)
di: de Carvalho, Marcus Vinicius Leal, et al.
Pubblicazione: (2024)
Governed Capability Evolution: Lifecycle-Time Compatibility Checking and Rollback for AI-Component-Based Systems, with Embodied Agents as Case Study
di: Qin, Xue, et al.
Pubblicazione: (2026)
di: Qin, Xue, et al.
Pubblicazione: (2026)
EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems
di: Qin, Xue, et al.
Pubblicazione: (2026)
di: Qin, Xue, et al.
Pubblicazione: (2026)
What Teaches Robots to Walk, Teaches Them to Trade too -- Regime Adaptive Execution using Informed Data and LLMs
di: Saqur, Raeid
Pubblicazione: (2024)
di: Saqur, Raeid
Pubblicazione: (2024)
Deployment-Time Reliability of Learned Robot Policies
di: Agia, Christopher
Pubblicazione: (2026)
di: Agia, Christopher
Pubblicazione: (2026)
AEROS: A Single-Agent Operating Architecture with Embodied Capability Modules
di: Qin, Xue, et al.
Pubblicazione: (2026)
di: Qin, Xue, et al.
Pubblicazione: (2026)
VLA Foundry: A Unified Framework for Training Vision-Language-Action Models
di: Mercat, Jean, et al.
Pubblicazione: (2026)
di: Mercat, Jean, et al.
Pubblicazione: (2026)
Working with Trouble and Failures in Conversation between Humans and Robots (WTF 2023) & Is CUI Design Ready Yet?
di: Förster, Frank, et al.
Pubblicazione: (2023)
di: Förster, Frank, et al.
Pubblicazione: (2023)
Beyond Greenfield: The D3 Framework for AI-Driven Productivity in Brownfield Engineering
di: Sharma, Krishna Kumaar
Pubblicazione: (2025)
di: Sharma, Krishna Kumaar
Pubblicazione: (2025)
Speech to Reality: On-Demand Production using Natural Language, 3D Generative AI, and Discrete Robotic Assembly
di: Kyaw, Alexander Htet, et al.
Pubblicazione: (2024)
di: Kyaw, Alexander Htet, et al.
Pubblicazione: (2024)
RoboScript: Code Generation for Free-Form Manipulation Tasks across Real and Simulation
di: Chen, Junting, et al.
Pubblicazione: (2024)
di: Chen, Junting, et al.
Pubblicazione: (2024)
PerspAct: Enhancing LLM Situated Collaboration Skills through Perspective Taking and Active Vision
di: Patania, Sabrina, et al.
Pubblicazione: (2025)
di: Patania, Sabrina, et al.
Pubblicazione: (2025)
A Replicable Robotics Awareness Method Using LLM-Enabled Robotics Interaction: Evidence from a Corporate Challenge
di: Prieto, S. A., et al.
Pubblicazione: (2026)
di: Prieto, S. A., et al.
Pubblicazione: (2026)
Conservative Bias in Multi-Teacher Learning: Why Agents Prefer Low-Reward Advisors
di: Mesto, Maher, et al.
Pubblicazione: (2025)
di: Mesto, Maher, et al.
Pubblicazione: (2025)
OWMM-Agent: Open World Mobile Manipulation With Multi-modal Agentic Data Synthesis
di: Chen, Junting, et al.
Pubblicazione: (2025)
di: Chen, Junting, et al.
Pubblicazione: (2025)
Cortex 2.0: Grounding World Models in Real-World Industrial Deployment
di: Aida, Adriana, et al.
Pubblicazione: (2026)
di: Aida, Adriana, et al.
Pubblicazione: (2026)
Robot guide with multi-agent control and automatic scenario generation with LLM
di: Moskovskaya, Elizaveta D., et al.
Pubblicazione: (2025)
di: Moskovskaya, Elizaveta D., et al.
Pubblicazione: (2025)
3D-Anchored Lookahead Planning for Persistent Robotic Scene Memory via World-Model-Based MCTS
di: Sidik, Bronislav, et al.
Pubblicazione: (2026)
di: Sidik, Bronislav, et al.
Pubblicazione: (2026)
Federated Single-Agent Robotics: Multi-Robot Coordination Without Intra-Robot Multi-Agent Fragmentation
di: Qin, Xue, et al.
Pubblicazione: (2026)
di: Qin, Xue, et al.
Pubblicazione: (2026)
Who Sees What? Structured Thought-Action Sequences for Epistemic Reasoning in LLMs
di: Annese, Luca, et al.
Pubblicazione: (2025)
di: Annese, Luca, et al.
Pubblicazione: (2025)
A Roadmap for Embodied and Social Grounding in LLMs
di: Incao, Sara, et al.
Pubblicazione: (2024)
di: Incao, Sara, et al.
Pubblicazione: (2024)
Learning Affordances at Inference-Time for Vision-Language-Action Models
di: Shah, Ameesh, et al.
Pubblicazione: (2025)
di: Shah, Ameesh, et al.
Pubblicazione: (2025)
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
Domain-Specific Fine-Tuning of Large Language Models for Interactive Robot Programming
di: Alt, Benjamin, et al.
Pubblicazione: (2023)
di: Alt, Benjamin, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Natural Language as Policies: Reasoning for Coordinate-Level Embodied Control with LLMs
di: Mikami, Yusuke, et al.
Pubblicazione: (2024) -
BRAVE: Brain-Controlled Prosthetic Arm with Voice Integration and Embodied Learning for Enhanced Mobility
di: Basit, Abdul, et al.
Pubblicazione: (2025) -
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025) -
LiloDriver: A Lifelong Learning Framework for Closed-loop Motion Planning in Long-tail Autonomous Driving Scenarios
di: Yao, Huaiyuan, et al.
Pubblicazione: (2025) -
MAP: Multi-user Personalization with Collaborative LLM-powered Agents
di: Lee, Christine, et al.
Pubblicazione: (2025)