RoboStressBench: Benchmarking VLM Robustness to Physical Visual Stress in Embodied Scenes
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Leyi, Zhao, Yifan, Zhang, Jinjie, Chen, Suzeyu, Chen, Wosong, Chen, Zhifei, Xu, Tianshuo, He, Qingchun, Hu, Hongxin, Huang, Haojian, Wei, Yangkai, Li, Wenqian, Li, Yinchuan, Chen, Ying-Cong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Right Inference Strategy Is All You Need: Nearly Training-Free Domain-Wise Inference for EgoCross Challenge
by: Wu, Leyi, et al.
Published: (2026)
by: Wu, Leyi, et al.
Published: (2026)
SPOT-Occ: Sparse Prototype-guided Transformer for Camera-based 3D Occupancy Prediction
by: Chen, Suzeyu, et al.
Published: (2026)
by: Chen, Suzeyu, et al.
Published: (2026)
Motion Forcing: A Decoupled Framework for Robust Video Generation in Motion Dynamics
by: Xu, Tianshuo, et al.
Published: (2026)
by: Xu, Tianshuo, et al.
Published: (2026)
Find, Fix, Reason: Context Repair for Video Reasoning
by: Huang, Haojian, et al.
Published: (2026)
by: Huang, Haojian, et al.
Published: (2026)
Affordance Agent Harness: Verification-Gated Skill Orchestration
by: Huang, Haojian, et al.
Published: (2026)
by: Huang, Haojian, et al.
Published: (2026)
UniCalli: A Unified Diffusion Framework for Column-Level Generation and Recognition of Chinese Calligraphy
by: Xu, Tianshuo, et al.
Published: (2025)
by: Xu, Tianshuo, et al.
Published: (2025)
PhysToolBench: Benchmarking Physical Tool Understanding for MLLMs
by: Zhang, Zixin, et al.
Published: (2025)
by: Zhang, Zixin, et al.
Published: (2025)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
by: Xu, Tianshuo, et al.
Published: (2024)
by: Xu, Tianshuo, et al.
Published: (2024)
Hydrochlorothiazide Improves Cardiac Remodeling in Heart Failure Rats by Reducing Oxidative Stress
by: Jinghong Luo, et al.
Published: (2024)
by: Jinghong Luo, et al.
Published: (2024)
RoboBlockly Studio: Conversational Block Programming with Embodied Robot Feedback for Computational Thinking
by: Li, Leyi, et al.
Published: (2026)
by: Li, Leyi, et al.
Published: (2026)
RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents
by: Yeke, Doguhuan, et al.
Published: (2026)
by: Yeke, Doguhuan, et al.
Published: (2026)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
by: Li, Huiqiong, et al.
Published: (2026)
by: Li, Huiqiong, et al.
Published: (2026)
RoboScape: Physics-informed Embodied World Model
by: Shang, Yu, et al.
Published: (2025)
by: Shang, Yu, et al.
Published: (2025)
AirCopBench: A Benchmark for Multi-drone Collaborative Embodied Perception and Reasoning
by: Zha, Jirong, et al.
Published: (2025)
by: Zha, Jirong, et al.
Published: (2025)
RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation
by: Jiang, Feng, et al.
Published: (2026)
by: Jiang, Feng, et al.
Published: (2026)
CleanUpBench: Embodied Sweeping and Grasping Benchmark
by: Li, Wenbo, et al.
Published: (2025)
by: Li, Wenbo, et al.
Published: (2025)
Global boundedness and asymptotic stability of the Keller-Segel system with logistic-type source in the whole space
by: Li, Qingchun, et al.
Published: (2024)
by: Li, Qingchun, et al.
Published: (2024)
EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems
by: Qin, Xue, et al.
Published: (2026)
by: Qin, Xue, et al.
Published: (2026)
STANCE: Motion Coherent Video Generation Via Sparse-to-Dense Anchored Encoding
by: Chen, Zhifei, et al.
Published: (2025)
by: Chen, Zhifei, et al.
Published: (2025)
RoboView-Bias: Benchmarking Visual Bias in Embodied Agents for Robotic Manipulation
by: Liu, Enguang, et al.
Published: (2025)
by: Liu, Enguang, et al.
Published: (2025)
RoboTidy : A 3D Gaussian Splatting Household Tidying Benchmark for Embodied Navigation and Action
by: Sun, Xiaoquan, et al.
Published: (2025)
by: Sun, Xiaoquan, et al.
Published: (2025)
IS-Bench: Evaluating Interactive Safety of VLM-Driven Embodied Agents in Daily Household Tasks
by: Lu, Xiaoya, et al.
Published: (2025)
by: Lu, Xiaoya, et al.
Published: (2025)
Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
by: Chen, Kaiyuan, et al.
Published: (2025)
by: Chen, Kaiyuan, et al.
Published: (2025)
CalliMaster: Mastering Page-level Chinese Calligraphy via Layout-guided Spatial Planning
by: Xu, Tianshuo, et al.
Published: (2026)
by: Xu, Tianshuo, et al.
Published: (2026)
3D-Mem: 3D Scene Memory for Embodied Exploration and Reasoning
by: Yang, Yuncong, et al.
Published: (2024)
by: Yang, Yuncong, et al.
Published: (2024)
PushupBench: Your VLM is not good at counting pushups
by: Li, Shengzhi, et al.
Published: (2026)
by: Li, Shengzhi, et al.
Published: (2026)
HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System
by: Yang, Tianshuo, et al.
Published: (2026)
by: Yang, Tianshuo, et al.
Published: (2026)
RoboLayout: Differentiable 3D Scene Generation for Embodied Agents
by: Shamsaddinlou, Ali
Published: (2026)
by: Shamsaddinlou, Ali
Published: (2026)
Uni-Renderer: Unifying Rendering and Inverse Rendering Via Dual Stream Diffusion
by: Chen, Zhifei, et al.
Published: (2024)
by: Chen, Zhifei, et al.
Published: (2024)
FlexPainter: Flexible and Multi-View Consistent Texture Generation
by: Yan, Dongyu, et al.
Published: (2025)
by: Yan, Dongyu, et al.
Published: (2025)
PolyBench: A Benchmark for Compositional Reasoning in Polyphonic Audio
by: Chen, Yuanjian, et al.
Published: (2026)
by: Chen, Yuanjian, et al.
Published: (2026)
Proximalized Preference Optimization for Diverse Feedback Types: A Decomposed Perspective on DPO
by: Guo, Kaiyang, et al.
Published: (2025)
by: Guo, Kaiyang, et al.
Published: (2025)
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models
by: Wu, Kui, et al.
Published: (2025)
by: Wu, Kui, et al.
Published: (2025)
VisualNeedle: Benchmarking Active Visual Search in Information-Dense Scenes
by: Chen, Jingru, et al.
Published: (2026)
by: Chen, Jingru, et al.
Published: (2026)
EvoEmpirBench: Dynamic Spatial Reasoning with Agent-ExpVer
by: Zhao, Pukun, et al.
Published: (2025)
by: Zhao, Pukun, et al.
Published: (2025)
RoboBenchMart: Benchmarking Robots in Retail Environment
by: Soshin, Konstantin, et al.
Published: (2025)
by: Soshin, Konstantin, et al.
Published: (2025)
Transcription Profiling of Potato Leaves in Response to Heat Stress at Single‐Cell Resolution
by: Shiqi Wen, et al.
Published: (2026)
by: Shiqi Wen, et al.
Published: (2026)
ROOT: VLM based System for Indoor Scene Understanding and Beyond
by: Wang, Yonghui, et al.
Published: (2024)
by: Wang, Yonghui, et al.
Published: (2024)
Color‐Resolved Stress Sensing Visualization Based on Ratiometric Mechanoluminescence: A Universal Strategy via Radiative Energy Transfer
by: Xiaoyang Liang, et al.
Published: (2026)
by: Xiaoyang Liang, et al.
Published: (2026)
UrbanVideo-Bench: Benchmarking Vision-Language Models on Embodied Intelligence with Video Data in Urban Spaces
by: Zhao, Baining, et al.
Published: (2025)
by: Zhao, Baining, et al.
Published: (2025)
Similar Items
-
The Right Inference Strategy Is All You Need: Nearly Training-Free Domain-Wise Inference for EgoCross Challenge
by: Wu, Leyi, et al.
Published: (2026) -
SPOT-Occ: Sparse Prototype-guided Transformer for Camera-based 3D Occupancy Prediction
by: Chen, Suzeyu, et al.
Published: (2026) -
Motion Forcing: A Decoupled Framework for Robust Video Generation in Motion Dynamics
by: Xu, Tianshuo, et al.
Published: (2026) -
Find, Fix, Reason: Context Repair for Video Reasoning
by: Huang, Haojian, et al.
Published: (2026) -
Affordance Agent Harness: Verification-Gated Skill Orchestration
by: Huang, Haojian, et al.
Published: (2026)