ReLIC: A Recipe for 64k Steps of In-Context Reinforcement Learning for Embodied AI
Fuente:
arXiv
Saved in:
| Main Authors: | Elawady, Ahmad, Chhablani, Gunjan, Ramrakhya, Ram, Yadav, Karmesh, Batra, Dhruv, Kira, Zsolt, Szot, Andrew |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GOAT-Bench: A Benchmark for Multi-Modal Lifelong Navigation
by: Khanna, Mukul, et al.
Published: (2024)
by: Khanna, Mukul, et al.
Published: (2024)
EmbodiedSplat: Personalized Real-to-Sim-to-Real Navigation with Gaussian Splats from a Mobile Device
by: Chhablani, Gunjan, et al.
Published: (2025)
by: Chhablani, Gunjan, et al.
Published: (2025)
Memo: Training Memory-Efficient Embodied Agents with Reinforcement Learning
by: Gupta, Gunshi, et al.
Published: (2025)
by: Gupta, Gunshi, et al.
Published: (2025)
Seeing the Unseen: Visual Common Sense for Semantic Placement
by: Ramrakhya, Ram, et al.
Published: (2024)
by: Ramrakhya, Ram, et al.
Published: (2024)
Grounding Multimodal LLMs to Embodied Agents that Ask for Help with Reinforcement Learning
by: Ramrakhya, Ram, et al.
Published: (2025)
by: Ramrakhya, Ram, et al.
Published: (2025)
FindingDory: A Benchmark to Evaluate Memory in Embodied Agents
by: Yadav, Karmesh, et al.
Published: (2025)
by: Yadav, Karmesh, et al.
Published: (2025)
ReLIC-SGG: Relation Lattice Completion for Open-Vocabulary Scene Graph Generation
by: Hosseini, Amir, et al.
Published: (2026)
by: Hosseini, Amir, et al.
Published: (2026)
Pre-trained Text-to-Image Diffusion Models Are Versatile Representation Learners for Control
by: Gupta, Gunshi, et al.
Published: (2024)
by: Gupta, Gunshi, et al.
Published: (2024)
Let's Think in Two Steps: Mitigating Agreement Bias in MLLMs with Self-Grounded Verification
by: Andrade, Moises, et al.
Published: (2025)
by: Andrade, Moises, et al.
Published: (2025)
HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation
by: Yokoyama, Naoki, et al.
Published: (2024)
by: Yokoyama, Naoki, et al.
Published: (2024)
Reinforcement Learning via Auxiliary Task Distillation
by: Harish, Abhinav Narayan, et al.
Published: (2024)
by: Harish, Abhinav Narayan, et al.
Published: (2024)
From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons
by: Szot, Andrew, et al.
Published: (2024)
by: Szot, Andrew, et al.
Published: (2024)
Grounding Multimodal Large Language Models in Actions
by: Szot, Andrew, et al.
Published: (2024)
by: Szot, Andrew, et al.
Published: (2024)
Scaling Synthetic Task Generation for Agents via Exploration
by: Ramrakhya, Ram, et al.
Published: (2025)
by: Ramrakhya, Ram, et al.
Published: (2025)
PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks
by: Chang, Matthew, et al.
Published: (2024)
by: Chang, Matthew, et al.
Published: (2024)
Where are we in the search for an Artificial Visual Cortex for Embodied Intelligence?
by: Majumdar, Arjun, et al.
Published: (2023)
by: Majumdar, Arjun, et al.
Published: (2023)
Towards Open-World Mobile Manipulation in Homes: Lessons from the Neurips 2023 HomeRobot Open Vocabulary Mobile Manipulation Challenge
by: Yenamandra, Sriram, et al.
Published: (2024)
by: Yenamandra, Sriram, et al.
Published: (2024)
SAGE: Sink-Aware Grounded Decoding for Multimodal Hallucination Mitigation
by: Shukla, Tripti, et al.
Published: (2026)
by: Shukla, Tripti, et al.
Published: (2026)
HomeRobot: Open-Vocabulary Mobile Manipulation
by: Yenamandra, Sriram, et al.
Published: (2023)
by: Yenamandra, Sriram, et al.
Published: (2023)
Hierarchical Reinforcement Learning and Value Optimization for Challenging Quadruped Locomotion
by: Coholich, Jeremiah, et al.
Published: (2025)
by: Coholich, Jeremiah, et al.
Published: (2025)
An Overview of Recent Developments on Electrodes Modified with Bacteriophages
by: Katarzyna Szot‐Karpińska
Published: (2025)
by: Katarzyna Szot‐Karpińska
Published: (2025)
LA TRANSICIÓN DEMOGRÁFICO-EPIDEMIOLÓGICA EN CHILE, 1960-2001
by: Jorge Szot Meza
Published: (2003)
by: Jorge Szot Meza
Published: (2003)
Large Language Models as Generalizable Policies for Embodied Tasks
by: Szot, Andrew, et al.
Published: (2023)
by: Szot, Andrew, et al.
Published: (2023)
UltraCUA: A Foundation Model for Computer Use Agents with Hybrid Action
by: Yang, Yuhao, et al.
Published: (2025)
by: Yang, Yuhao, et al.
Published: (2025)
What do we learn from a large-scale study of pre-trained visual representations in sim and real environments?
by: Silwal, Sneha, et al.
Published: (2023)
by: Silwal, Sneha, et al.
Published: (2023)
RecipeGen: A Step-Aligned Multimodal Benchmark for Real-World Recipe Generation
by: Zhang, Ruoxuan, et al.
Published: (2025)
by: Zhang, Ruoxuan, et al.
Published: (2025)
Rethinking Weight Decay for Robust Fine-Tuning of Foundation Models
by: Tian, Junjiao, et al.
Published: (2024)
by: Tian, Junjiao, et al.
Published: (2024)
Barrier Function Overrides For Non-Convex Fixed Wing Flight Control and Self-Driving Cars
by: Squires, Eric, et al.
Published: (2025)
by: Squires, Eric, et al.
Published: (2025)
QCrypton: A Unified Platform for AI/LLM Threat Detection and Post-Quantum Cryptographic Security Assessment
by: Jain, Gunjan
Published: (2026)
by: Jain, Gunjan
Published: (2026)
The MUG-10 Framework for Preventing Usability Issues in Mobile Application Development
by: Weichbroth, Pawel, et al.
Published: (2025)
by: Weichbroth, Pawel, et al.
Published: (2025)
Thousand-GPU Large-Scale Training and Optimization Recipe for AI-Native Cloud Embodied Intelligence Infrastructure
by: Guo, Yongjian, et al.
Published: (2026)
by: Guo, Yongjian, et al.
Published: (2026)
LongRecipe: Recipe for Efficient Long Context Generalization in Large Language Models
by: Hu, Zhiyuan, et al.
Published: (2024)
by: Hu, Zhiyuan, et al.
Published: (2024)
Contextual Self-paced Learning for Weakly Supervised Spatio-Temporal Video Grounding
by: Kumar, Akash, et al.
Published: (2025)
by: Kumar, Akash, et al.
Published: (2025)
N-QR: Natural Quick Response Codes for Multi-Robot Instance Correspondence
by: Glaser, Nathaniel Moore, et al.
Published: (2024)
by: Glaser, Nathaniel Moore, et al.
Published: (2024)
Expanding LLM Agent Boundaries with Strategy-Guided Exploration
by: Szot, Andrew, et al.
Published: (2026)
by: Szot, Andrew, et al.
Published: (2026)
TinyLIC-High efficiency lossy image compression method
by: Ma, Gaocheng, et al.
Published: (2024)
by: Ma, Gaocheng, et al.
Published: (2024)
Re/Embodied Data
by: Conradi, Florian, et al.
Published: (2026)
by: Conradi, Florian, et al.
Published: (2026)
Performance Analysis of a Lightweight Image Re‐Encryption Scheme for 5G HetNets
by: Ajay Kakkar, et al.
Published: (2025)
by: Ajay Kakkar, et al.
Published: (2025)
Mimicking or Reasoning: Rethinking Multi-Modal In-Context Learning in Vision-Language Models
by: Huang, Chengyue, et al.
Published: (2025)
by: Huang, Chengyue, et al.
Published: (2025)
Step Rejection Fine-Tuning: A Practical Distillation Recipe
by: Slinko, Igor, et al.
Published: (2026)
by: Slinko, Igor, et al.
Published: (2026)
Similar Items
-
GOAT-Bench: A Benchmark for Multi-Modal Lifelong Navigation
by: Khanna, Mukul, et al.
Published: (2024) -
EmbodiedSplat: Personalized Real-to-Sim-to-Real Navigation with Gaussian Splats from a Mobile Device
by: Chhablani, Gunjan, et al.
Published: (2025) -
Memo: Training Memory-Efficient Embodied Agents with Reinforcement Learning
by: Gupta, Gunshi, et al.
Published: (2025) -
Seeing the Unseen: Visual Common Sense for Semantic Placement
by: Ramrakhya, Ram, et al.
Published: (2024) -
Grounding Multimodal LLMs to Embodied Agents that Ask for Help with Reinforcement Learning
by: Ramrakhya, Ram, et al.
Published: (2025)