GOAT-Bench: A Benchmark for Multi-Modal Lifelong Navigation
Fuente:
arXiv
Salvato in:
| Autori principali: | Khanna, Mukul, Ramrakhya, Ram, Chhablani, Gunjan, Yenamandra, Sriram, Gervet, Theophile, Chang, Matthew, Kira, Zsolt, Chaplot, Devendra Singh, Batra, Dhruv, Mottaghi, Roozbeh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ReLIC: A Recipe for 64k Steps of In-Context Reinforcement Learning for Embodied AI
di: Elawady, Ahmad, et al.
Pubblicazione: (2024)
di: Elawady, Ahmad, et al.
Pubblicazione: (2024)
HomeRobot: Open-Vocabulary Mobile Manipulation
di: Yenamandra, Sriram, et al.
Pubblicazione: (2023)
di: Yenamandra, Sriram, et al.
Pubblicazione: (2023)
Grounding Multimodal LLMs to Embodied Agents that Ask for Help with Reinforcement Learning
di: Ramrakhya, Ram, et al.
Pubblicazione: (2025)
di: Ramrakhya, Ram, et al.
Pubblicazione: (2025)
EmbodiedSplat: Personalized Real-to-Sim-to-Real Navigation with Gaussian Splats from a Mobile Device
di: Chhablani, Gunjan, et al.
Pubblicazione: (2025)
di: Chhablani, Gunjan, et al.
Pubblicazione: (2025)
HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation
di: Yokoyama, Naoki, et al.
Pubblicazione: (2024)
di: Yokoyama, Naoki, et al.
Pubblicazione: (2024)
Seeing the Unseen: Visual Common Sense for Semantic Placement
di: Ramrakhya, Ram, et al.
Pubblicazione: (2024)
di: Ramrakhya, Ram, et al.
Pubblicazione: (2024)
Towards Open-World Mobile Manipulation in Homes: Lessons from the Neurips 2023 HomeRobot Open Vocabulary Mobile Manipulation Challenge
di: Yenamandra, Sriram, et al.
Pubblicazione: (2024)
di: Yenamandra, Sriram, et al.
Pubblicazione: (2024)
Situated Instruction Following
di: Min, So Yeon, et al.
Pubblicazione: (2024)
di: Min, So Yeon, et al.
Pubblicazione: (2024)
PARTNR: A Benchmark for Planning and Reasoning in Embodied Multi-agent Tasks
di: Chang, Matthew, et al.
Pubblicazione: (2024)
di: Chang, Matthew, et al.
Pubblicazione: (2024)
FRAMES-VQA: Benchmarking Fine-Tuning Robustness across Multi-Modal Shifts in Visual Question Answering
di: Huang, Chengyue, et al.
Pubblicazione: (2025)
di: Huang, Chengyue, et al.
Pubblicazione: (2025)
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
di: Dalal, Murtaza, et al.
Pubblicazione: (2024)
di: Dalal, Murtaza, et al.
Pubblicazione: (2024)
Track2Act: Predicting Point Tracks from Internet Videos enables Generalizable Robot Manipulation
di: Bharadhwaj, Homanga, et al.
Pubblicazione: (2024)
di: Bharadhwaj, Homanga, et al.
Pubblicazione: (2024)
ObjectForesight: Predicting Future 3D Object Trajectories from Human Videos
di: Soraki, Rustin, et al.
Pubblicazione: (2026)
di: Soraki, Rustin, et al.
Pubblicazione: (2026)
Pre-trained Text-to-Image Diffusion Models Are Versatile Representation Learners for Control
di: Gupta, Gunshi, et al.
Pubblicazione: (2024)
di: Gupta, Gunshi, et al.
Pubblicazione: (2024)
SOC Increase in UK Topsoils Is Most Likely due to SOC Vertical Redistribution: Comment on Bentley et al. (2025)
di: Vincent Chaplot
Pubblicazione: (2026)
di: Vincent Chaplot
Pubblicazione: (2026)
Degradable Strongly Entanglement Breaking Maps
di: Devendra, Repana, et al.
Pubblicazione: (2023)
di: Devendra, Repana, et al.
Pubblicazione: (2023)
GOAT-Bench: Safety Insights to Large Multimodal Models through Meme-Based Social Abuse
di: Lin, Hongzhan, et al.
Pubblicazione: (2024)
di: Lin, Hongzhan, et al.
Pubblicazione: (2024)
Contextual Self-paced Learning for Weakly Supervised Spatio-Temporal Video Grounding
di: Kumar, Akash, et al.
Pubblicazione: (2025)
di: Kumar, Akash, et al.
Pubblicazione: (2025)
SAGE: Sink-Aware Grounded Decoding for Multimodal Hallucination Mitigation
di: Shukla, Tripti, et al.
Pubblicazione: (2026)
di: Shukla, Tripti, et al.
Pubblicazione: (2026)
Effects of shade on guinea grass genotypes Megathyrsus maximus (Poales: Poaceae)
di: Devendra Ram Malaviya
Pubblicazione: (2020)
di: Devendra Ram Malaviya
Pubblicazione: (2020)
Cube Bench: A Benchmark for Spatial Visual Reasoning in MLLMs
di: Anand, Dhruv, et al.
Pubblicazione: (2025)
di: Anand, Dhruv, et al.
Pubblicazione: (2025)
Environmental geopolitics of the Caspian basin energy interactions
di: Mottaghi, A.
Pubblicazione: (2016)
di: Mottaghi, A.
Pubblicazione: (2016)
KnowMe-Bench: Benchmarking Person Understanding for Lifelong Digital Companions
di: Wu, Tingyu, et al.
Pubblicazione: (2026)
di: Wu, Tingyu, et al.
Pubblicazione: (2026)
Benchmarking Active Learning for NILM
di: Patel, Dhruv, et al.
Pubblicazione: (2024)
di: Patel, Dhruv, et al.
Pubblicazione: (2024)
LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners
di: Zheng, Junhao, et al.
Pubblicazione: (2025)
di: Zheng, Junhao, et al.
Pubblicazione: (2025)
Benchmarking Humans and Machines on Complex Multilingual Speech Understanding Tasks
di: Kankanala, Sai Samrat, et al.
Pubblicazione: (2025)
di: Kankanala, Sai Samrat, et al.
Pubblicazione: (2025)
FindingDory: A Benchmark to Evaluate Memory in Embodied Agents
di: Yadav, Karmesh, et al.
Pubblicazione: (2025)
di: Yadav, Karmesh, et al.
Pubblicazione: (2025)
Cover crop studies: The need for more reliable data
di: Vincent Chaplot, et al.
Pubblicazione: (2024)
di: Vincent Chaplot, et al.
Pubblicazione: (2024)
Controllable Human-Object Interaction Synthesis
di: Li, Jiaman, et al.
Pubblicazione: (2023)
di: Li, Jiaman, et al.
Pubblicazione: (2023)
Lifelong Embodied Navigation Learning
di: Wang, Xudong, et al.
Pubblicazione: (2026)
di: Wang, Xudong, et al.
Pubblicazione: (2026)
Rethinking Weight Decay for Robust Fine-Tuning of Foundation Models
di: Tian, Junjiao, et al.
Pubblicazione: (2024)
di: Tian, Junjiao, et al.
Pubblicazione: (2024)
Barrier Function Overrides For Non-Convex Fixed Wing Flight Control and Self-Driving Cars
di: Squires, Eric, et al.
Pubblicazione: (2025)
di: Squires, Eric, et al.
Pubblicazione: (2025)
GOAT: A Training Framework for Goal-Oriented Agent with Tools
di: Min, Hyunji, et al.
Pubblicazione: (2025)
di: Min, Hyunji, et al.
Pubblicazione: (2025)
Automated Red Teaming with GOAT: the Generative Offensive Agent Tester
di: Pavlova, Maya, et al.
Pubblicazione: (2024)
di: Pavlova, Maya, et al.
Pubblicazione: (2024)
GOAT: A Global Optimization Algorithm for Molecules and Atomic Clusters
di: Bernardo de Souza
Pubblicazione: (2025)
di: Bernardo de Souza
Pubblicazione: (2025)
GOAT: A Global Optimization Algorithm for Molecules and Atomic Clusters
di: Bernardo de Souza
Pubblicazione: (2025)
di: Bernardo de Souza
Pubblicazione: (2025)
Light‐Driven Micromotors for Microscale Metal Ion Sensing Applications in Fluid Medium
di: Srikanta Debata, et al.
Pubblicazione: (2025)
di: Srikanta Debata, et al.
Pubblicazione: (2025)
SafeManip: A Property-Driven Benchmark for Temporal Safety Evaluation in Robotic Manipulation
di: Huang, Chengyue, et al.
Pubblicazione: (2026)
di: Huang, Chengyue, et al.
Pubblicazione: (2026)
MCJudgeBench: A Benchmark for Constraint-Level Judge Evaluation in Multi-Constraint Instruction Following
di: Lee, Jaeyun, et al.
Pubblicazione: (2026)
di: Lee, Jaeyun, et al.
Pubblicazione: (2026)
ADAPT: Actively Discovering and Adapting to Preferences for any Task
di: Patel, Maithili, et al.
Pubblicazione: (2025)
di: Patel, Maithili, et al.
Pubblicazione: (2025)
Documenti analoghi
-
ReLIC: A Recipe for 64k Steps of In-Context Reinforcement Learning for Embodied AI
di: Elawady, Ahmad, et al.
Pubblicazione: (2024) -
HomeRobot: Open-Vocabulary Mobile Manipulation
di: Yenamandra, Sriram, et al.
Pubblicazione: (2023) -
Grounding Multimodal LLMs to Embodied Agents that Ask for Help with Reinforcement Learning
di: Ramrakhya, Ram, et al.
Pubblicazione: (2025) -
EmbodiedSplat: Personalized Real-to-Sim-to-Real Navigation with Gaussian Splats from a Mobile Device
di: Chhablani, Gunjan, et al.
Pubblicazione: (2025) -
HM3D-OVON: A Dataset and Benchmark for Open-Vocabulary Object Goal Navigation
di: Yokoyama, Naoki, et al.
Pubblicazione: (2024)