RoboStressBench: Benchmarking VLM Robustness to Physical Visual Stress in Embodied Scenes
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wu, Leyi, Zhao, Yifan, Zhang, Jinjie, Chen, Suzeyu, Chen, Wosong, Chen, Zhifei, Xu, Tianshuo, He, Qingchun, Hu, Hongxin, Huang, Haojian, Wei, Yangkai, Li, Wenqian, Li, Yinchuan, Chen, Ying-Cong |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The Right Inference Strategy Is All You Need: Nearly Training-Free Domain-Wise Inference for EgoCross Challenge
par: Wu, Leyi, et autres
Publié: (2026)
par: Wu, Leyi, et autres
Publié: (2026)
SPOT-Occ: Sparse Prototype-guided Transformer for Camera-based 3D Occupancy Prediction
par: Chen, Suzeyu, et autres
Publié: (2026)
par: Chen, Suzeyu, et autres
Publié: (2026)
Motion Forcing: A Decoupled Framework for Robust Video Generation in Motion Dynamics
par: Xu, Tianshuo, et autres
Publié: (2026)
par: Xu, Tianshuo, et autres
Publié: (2026)
Find, Fix, Reason: Context Repair for Video Reasoning
par: Huang, Haojian, et autres
Publié: (2026)
par: Huang, Haojian, et autres
Publié: (2026)
Affordance Agent Harness: Verification-Gated Skill Orchestration
par: Huang, Haojian, et autres
Publié: (2026)
par: Huang, Haojian, et autres
Publié: (2026)
UniCalli: A Unified Diffusion Framework for Column-Level Generation and Recognition of Chinese Calligraphy
par: Xu, Tianshuo, et autres
Publié: (2025)
par: Xu, Tianshuo, et autres
Publié: (2025)
PhysToolBench: Benchmarking Physical Tool Understanding for MLLMs
par: Zhang, Zixin, et autres
Publié: (2025)
par: Zhang, Zixin, et autres
Publié: (2025)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
par: Xu, Tianshuo, et autres
Publié: (2024)
par: Xu, Tianshuo, et autres
Publié: (2024)
Hydrochlorothiazide Improves Cardiac Remodeling in Heart Failure Rats by Reducing Oxidative Stress
par: Jinghong Luo, et autres
Publié: (2024)
par: Jinghong Luo, et autres
Publié: (2024)
RoboBlockly Studio: Conversational Block Programming with Embodied Robot Feedback for Computational Thinking
par: Li, Leyi, et autres
Publié: (2026)
par: Li, Leyi, et autres
Publié: (2026)
RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents
par: Yeke, Doguhuan, et autres
Publié: (2026)
par: Yeke, Doguhuan, et autres
Publié: (2026)
RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
par: Li, Huiqiong, et autres
Publié: (2026)
par: Li, Huiqiong, et autres
Publié: (2026)
RoboScape: Physics-informed Embodied World Model
par: Shang, Yu, et autres
Publié: (2025)
par: Shang, Yu, et autres
Publié: (2025)
AirCopBench: A Benchmark for Multi-drone Collaborative Embodied Perception and Reasoning
par: Zha, Jirong, et autres
Publié: (2025)
par: Zha, Jirong, et autres
Publié: (2025)
RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation
par: Jiang, Feng, et autres
Publié: (2026)
par: Jiang, Feng, et autres
Publié: (2026)
CleanUpBench: Embodied Sweeping and Grasping Benchmark
par: Li, Wenbo, et autres
Publié: (2025)
par: Li, Wenbo, et autres
Publié: (2025)
Global boundedness and asymptotic stability of the Keller-Segel system with logistic-type source in the whole space
par: Li, Qingchun, et autres
Publié: (2024)
par: Li, Qingchun, et autres
Publié: (2024)
EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems
par: Qin, Xue, et autres
Publié: (2026)
par: Qin, Xue, et autres
Publié: (2026)
STANCE: Motion Coherent Video Generation Via Sparse-to-Dense Anchored Encoding
par: Chen, Zhifei, et autres
Publié: (2025)
par: Chen, Zhifei, et autres
Publié: (2025)
RoboView-Bias: Benchmarking Visual Bias in Embodied Agents for Robotic Manipulation
par: Liu, Enguang, et autres
Publié: (2025)
par: Liu, Enguang, et autres
Publié: (2025)
RoboTidy : A 3D Gaussian Splatting Household Tidying Benchmark for Embodied Navigation and Action
par: Sun, Xiaoquan, et autres
Publié: (2025)
par: Sun, Xiaoquan, et autres
Publié: (2025)
IS-Bench: Evaluating Interactive Safety of VLM-Driven Embodied Agents in Daily Household Tasks
par: Lu, Xiaoya, et autres
Publié: (2025)
par: Lu, Xiaoya, et autres
Publié: (2025)
Robo2VLM: Visual Question Answering from Large-Scale In-the-Wild Robot Manipulation Datasets
par: Chen, Kaiyuan, et autres
Publié: (2025)
par: Chen, Kaiyuan, et autres
Publié: (2025)
CalliMaster: Mastering Page-level Chinese Calligraphy via Layout-guided Spatial Planning
par: Xu, Tianshuo, et autres
Publié: (2026)
par: Xu, Tianshuo, et autres
Publié: (2026)
3D-Mem: 3D Scene Memory for Embodied Exploration and Reasoning
par: Yang, Yuncong, et autres
Publié: (2024)
par: Yang, Yuncong, et autres
Publié: (2024)
PushupBench: Your VLM is not good at counting pushups
par: Li, Shengzhi, et autres
Publié: (2026)
par: Li, Shengzhi, et autres
Publié: (2026)
HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System
par: Yang, Tianshuo, et autres
Publié: (2026)
par: Yang, Tianshuo, et autres
Publié: (2026)
RoboLayout: Differentiable 3D Scene Generation for Embodied Agents
par: Shamsaddinlou, Ali
Publié: (2026)
par: Shamsaddinlou, Ali
Publié: (2026)
Uni-Renderer: Unifying Rendering and Inverse Rendering Via Dual Stream Diffusion
par: Chen, Zhifei, et autres
Publié: (2024)
par: Chen, Zhifei, et autres
Publié: (2024)
FlexPainter: Flexible and Multi-View Consistent Texture Generation
par: Yan, Dongyu, et autres
Publié: (2025)
par: Yan, Dongyu, et autres
Publié: (2025)
PolyBench: A Benchmark for Compositional Reasoning in Polyphonic Audio
par: Chen, Yuanjian, et autres
Publié: (2026)
par: Chen, Yuanjian, et autres
Publié: (2026)
Proximalized Preference Optimization for Diverse Feedback Types: A Decomposed Perspective on DPO
par: Guo, Kaiyang, et autres
Publié: (2025)
par: Guo, Kaiyang, et autres
Publié: (2025)
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models
par: Wu, Kui, et autres
Publié: (2025)
par: Wu, Kui, et autres
Publié: (2025)
VisualNeedle: Benchmarking Active Visual Search in Information-Dense Scenes
par: Chen, Jingru, et autres
Publié: (2026)
par: Chen, Jingru, et autres
Publié: (2026)
EvoEmpirBench: Dynamic Spatial Reasoning with Agent-ExpVer
par: Zhao, Pukun, et autres
Publié: (2025)
par: Zhao, Pukun, et autres
Publié: (2025)
RoboBenchMart: Benchmarking Robots in Retail Environment
par: Soshin, Konstantin, et autres
Publié: (2025)
par: Soshin, Konstantin, et autres
Publié: (2025)
Transcription Profiling of Potato Leaves in Response to Heat Stress at Single‐Cell Resolution
par: Shiqi Wen, et autres
Publié: (2026)
par: Shiqi Wen, et autres
Publié: (2026)
ROOT: VLM based System for Indoor Scene Understanding and Beyond
par: Wang, Yonghui, et autres
Publié: (2024)
par: Wang, Yonghui, et autres
Publié: (2024)
Color‐Resolved Stress Sensing Visualization Based on Ratiometric Mechanoluminescence: A Universal Strategy via Radiative Energy Transfer
par: Xiaoyang Liang, et autres
Publié: (2026)
par: Xiaoyang Liang, et autres
Publié: (2026)
UrbanVideo-Bench: Benchmarking Vision-Language Models on Embodied Intelligence with Video Data in Urban Spaces
par: Zhao, Baining, et autres
Publié: (2025)
par: Zhao, Baining, et autres
Publié: (2025)
Documents similaires
-
The Right Inference Strategy Is All You Need: Nearly Training-Free Domain-Wise Inference for EgoCross Challenge
par: Wu, Leyi, et autres
Publié: (2026) -
SPOT-Occ: Sparse Prototype-guided Transformer for Camera-based 3D Occupancy Prediction
par: Chen, Suzeyu, et autres
Publié: (2026) -
Motion Forcing: A Decoupled Framework for Robust Video Generation in Motion Dynamics
par: Xu, Tianshuo, et autres
Publié: (2026) -
Find, Fix, Reason: Context Repair for Video Reasoning
par: Huang, Haojian, et autres
Publié: (2026) -
Affordance Agent Harness: Verification-Gated Skill Orchestration
par: Huang, Haojian, et autres
Publié: (2026)