BOSS: Benchmark for Observation Space Shift in Long-Horizon Task
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Yue, Zhao, Linfeng, Ding, Mingyu, Bertasius, Gedas, Szafir, Daniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LiLo-VLA: Compositional Long-Horizon Manipulation via Linked Object-Centric Policies
von: Yang, Yue, et al.
Veröffentlicht: (2026)
von: Yang, Yue, et al.
Veröffentlicht: (2026)
ReBot: Scaling Robot Learning with Real-to-Sim-to-Real Robotic Video Synthesis
von: Fang, Yu, et al.
Veröffentlicht: (2025)
von: Fang, Yu, et al.
Veröffentlicht: (2025)
EgoMemReason: A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Video Understanding
von: Wang, Ziyang, et al.
Veröffentlicht: (2026)
von: Wang, Ziyang, et al.
Veröffentlicht: (2026)
World-Ego Modeling for Long-Horizon Evolution in Hybrid Embodied Tasks
von: Lin, Zuyao, et al.
Veröffentlicht: (2026)
von: Lin, Zuyao, et al.
Veröffentlicht: (2026)
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
von: Yuan, Puzhen, et al.
Veröffentlicht: (2025)
von: Yuan, Puzhen, et al.
Veröffentlicht: (2025)
VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks
von: Zhang, Shiduo, et al.
Veröffentlicht: (2024)
von: Zhang, Shiduo, et al.
Veröffentlicht: (2024)
Mixture of Horizons in Action Chunking
von: Jing, Dong, et al.
Veröffentlicht: (2025)
von: Jing, Dong, et al.
Veröffentlicht: (2025)
Chameleon: Episodic Memory for Long-Horizon Robotic Manipulation
von: Guo, Xinying, et al.
Veröffentlicht: (2026)
von: Guo, Xinying, et al.
Veröffentlicht: (2026)
TimeBlind: A Spatio-Temporal Compositionality Benchmark for Video LLMs
von: Li, Baiqi, et al.
Veröffentlicht: (2026)
von: Li, Baiqi, et al.
Veröffentlicht: (2026)
A Backbone for Long-Horizon Robot Task Understanding
von: Chen, Xiaoshuai, et al.
Veröffentlicht: (2024)
von: Chen, Xiaoshuai, et al.
Veröffentlicht: (2024)
ARM: Advantage Reward Modeling for Long-Horizon Manipulation
von: Mao, Yiming, et al.
Veröffentlicht: (2026)
von: Mao, Yiming, et al.
Veröffentlicht: (2026)
ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs
von: Wu, Xin, et al.
Veröffentlicht: (2026)
von: Wu, Xin, et al.
Veröffentlicht: (2026)
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
ClothPPO: A Proximal Policy Optimization Enhancing Framework for Robotic Cloth Manipulation with Observation-Aligned Action Spaces
von: Yang, Libing, et al.
Veröffentlicht: (2024)
von: Yang, Libing, et al.
Veröffentlicht: (2024)
TaPD: Temporal-adaptive Progressive Distillation for Observation-Adaptive Trajectory Forecasting in Autonomous Driving
von: Fan, Mingyu, et al.
Veröffentlicht: (2026)
von: Fan, Mingyu, et al.
Veröffentlicht: (2026)
NovaPlan: Zero-Shot Long-Horizon Manipulation via Closed-Loop Video Language Planning
von: Fu, Jiahui, et al.
Veröffentlicht: (2026)
von: Fu, Jiahui, et al.
Veröffentlicht: (2026)
SLIM: Sim-to-Real Legged Instructive Manipulation via Long-Horizon Visuomotor Learning
von: Zhang, Haichao, et al.
Veröffentlicht: (2025)
von: Zhang, Haichao, et al.
Veröffentlicht: (2025)
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs
von: Fang, Yu, et al.
Veröffentlicht: (2026)
von: Fang, Yu, et al.
Veröffentlicht: (2026)
VideoTree: Adaptive Tree-based Video Representation for LLM Reasoning on Long Videos
von: Wang, Ziyang, et al.
Veröffentlicht: (2024)
von: Wang, Ziyang, et al.
Veröffentlicht: (2024)
LaMMA-P: Generalizable Multi-Agent Long-Horizon Task Allocation and Planning with LM-Driven PDDL Planner
von: Zhang, Xiaopan, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaopan, et al.
Veröffentlicht: (2024)
RDD: Retrieval-Based Demonstration Decomposer for Planner Alignment in Long-Horizon Tasks
von: Yan, Mingxuan, et al.
Veröffentlicht: (2025)
von: Yan, Mingxuan, et al.
Veröffentlicht: (2025)
SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios
von: Lin, Jieru, et al.
Veröffentlicht: (2025)
von: Lin, Jieru, et al.
Veröffentlicht: (2025)
ODYSSEY: Open-World Quadrupeds Exploration and Manipulation for Long-Horizon Tasks
von: Wang, Kaijun, et al.
Veröffentlicht: (2025)
von: Wang, Kaijun, et al.
Veröffentlicht: (2025)
SPIRE: Synergistic Planning, Imitation, and Reinforcement Learning for Long-Horizon Manipulation
von: Zhou, Zihan, et al.
Veröffentlicht: (2024)
von: Zhou, Zihan, et al.
Veröffentlicht: (2024)
Grounding 3D Object Affordance with Language Instructions, Visual Observations and Interactions
von: Zhu, He, et al.
Veröffentlicht: (2025)
von: Zhu, He, et al.
Veröffentlicht: (2025)
MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation
von: Zhang, Pingrui, et al.
Veröffentlicht: (2025)
von: Zhang, Pingrui, et al.
Veröffentlicht: (2025)
Accelerating Transformer-Based Monocular SLAM via Geometric Utility Scoring
von: Xiong, Xinmiao, et al.
Veröffentlicht: (2026)
von: Xiong, Xinmiao, et al.
Veröffentlicht: (2026)
Video-RTS: Rethinking Reinforcement Learning and Test-Time Scaling for Efficient and Enhanced Video Reasoning
von: Wang, Ziyang, et al.
Veröffentlicht: (2025)
von: Wang, Ziyang, et al.
Veröffentlicht: (2025)
PhyGrasp: Generalizing Robotic Grasping with Physics-informed Large Multimodal Models
von: Guo, Dingkun, et al.
Veröffentlicht: (2024)
von: Guo, Dingkun, et al.
Veröffentlicht: (2024)
Data Shift of Object Detection in Autonomous Driving
von: Xu, Lida
Veröffentlicht: (2025)
von: Xu, Lida
Veröffentlicht: (2025)
Integrating Object Detection Modality into Visual Language Model for Enhanced Autonomous Driving Agent
von: He, Linfeng, et al.
Veröffentlicht: (2024)
von: He, Linfeng, et al.
Veröffentlicht: (2024)
TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning
von: Feng, ZhiYuan, et al.
Veröffentlicht: (2026)
von: Feng, ZhiYuan, et al.
Veröffentlicht: (2026)
REST: Receding Horizon Explorative Steiner Tree for Zero-Shot Object-Goal Navigation
von: Xiao, Shuqi, et al.
Veröffentlicht: (2026)
von: Xiao, Shuqi, et al.
Veröffentlicht: (2026)
MSC-Bench: Benchmarking and Analyzing Multi-Sensor Corruption for Driving Perception
von: Hao, Xiaoshuai, et al.
Veröffentlicht: (2025)
von: Hao, Xiaoshuai, et al.
Veröffentlicht: (2025)
Conditioning Latent-Space Clusters for Real-World Anomaly Classification
von: Bogdoll, Daniel, et al.
Veröffentlicht: (2023)
von: Bogdoll, Daniel, et al.
Veröffentlicht: (2023)
Towards Long-horizon Embodied Agents with Tool-Aligned Vision-Language-Action Models
von: Lei, Zixing, et al.
Veröffentlicht: (2026)
von: Lei, Zixing, et al.
Veröffentlicht: (2026)
TransFusionOdom: Interpretable Transformer-based LiDAR-Inertial Fusion Odometry Estimation
von: Sun, Leyuan, et al.
Veröffentlicht: (2023)
von: Sun, Leyuan, et al.
Veröffentlicht: (2023)
Long-term Traffic Simulation with Interleaved Autoregressive Motion and Scenario Generation
von: Yang, Xiuyu, et al.
Veröffentlicht: (2025)
von: Yang, Xiuyu, et al.
Veröffentlicht: (2025)
OccSim: Multi-kilometer Simulation with Long-horizon Occupancy World Models
von: Liu, Tianran, et al.
Veröffentlicht: (2026)
von: Liu, Tianran, et al.
Veröffentlicht: (2026)
OpenObj: Open-Vocabulary Object-Level Neural Radiance Fields with Fine-Grained Understanding
von: Deng, Yinan, et al.
Veröffentlicht: (2024)
von: Deng, Yinan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
LiLo-VLA: Compositional Long-Horizon Manipulation via Linked Object-Centric Policies
von: Yang, Yue, et al.
Veröffentlicht: (2026) -
ReBot: Scaling Robot Learning with Real-to-Sim-to-Real Robotic Video Synthesis
von: Fang, Yu, et al.
Veröffentlicht: (2025) -
EgoMemReason: A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Video Understanding
von: Wang, Ziyang, et al.
Veröffentlicht: (2026) -
World-Ego Modeling for Long-Horizon Evolution in Hybrid Embodied Tasks
von: Lin, Zuyao, et al.
Veröffentlicht: (2026) -
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
von: Yuan, Puzhen, et al.
Veröffentlicht: (2025)