ObjChangeVR: Object State Change Reasoning from Continuous Egocentric Views in VR Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ding, Shiyi, Wu, Shaoen, Chen, Ying |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RAG-VR: Leveraging Retrieval-Augmented Generation for 3D Question Answering in VR Environments
von: Ding, Shiyi, et al.
Veröffentlicht: (2025)
von: Ding, Shiyi, et al.
Veröffentlicht: (2025)
ObjFormer: Learning Land-Cover Changes From Paired OSM Data and Optical High-Resolution Imagery via Object-Guided Transformer
von: Chen, Hongruixuan, et al.
Veröffentlicht: (2023)
von: Chen, Hongruixuan, et al.
Veröffentlicht: (2023)
MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning
von: Jiang, Zheng, et al.
Veröffentlicht: (2026)
von: Jiang, Zheng, et al.
Veröffentlicht: (2026)
Fast Registration of Photorealistic Avatars for VR Facial Animation
von: Patel, Chaitanya, et al.
Veröffentlicht: (2024)
von: Patel, Chaitanya, et al.
Veröffentlicht: (2024)
OpenObj: Open-Vocabulary Object-Level Neural Radiance Fields with Fine-Grained Understanding
von: Deng, Yinan, et al.
Veröffentlicht: (2024)
von: Deng, Yinan, et al.
Veröffentlicht: (2024)
OSCBench: Benchmarking Object State Change in Text-to-Video Generation
von: Han, Xianjing, et al.
Veröffentlicht: (2026)
von: Han, Xianjing, et al.
Veröffentlicht: (2026)
VOODOO XP: Expressive One-Shot Head Reenactment for VR Telepresence
von: Tran, Phong, et al.
Veröffentlicht: (2024)
von: Tran, Phong, et al.
Veröffentlicht: (2024)
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
von: Banerjee, Prithviraj, et al.
Veröffentlicht: (2024)
von: Banerjee, Prithviraj, et al.
Veröffentlicht: (2024)
Aria-NeRF: Multimodal Egocentric View Synthesis
von: Sun, Jiankai, et al.
Veröffentlicht: (2023)
von: Sun, Jiankai, et al.
Veröffentlicht: (2023)
Object Aware Egocentric Online Action Detection
von: An, Joungbin, et al.
Veröffentlicht: (2024)
von: An, Joungbin, et al.
Veröffentlicht: (2024)
FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentation
von: Zhang, Zihui, et al.
Veröffentlicht: (2026)
von: Zhang, Zihui, et al.
Veröffentlicht: (2026)
EgoWorld: Translating Exocentric View to Egocentric View using Rich Exocentric Observations
von: Park, Junho, et al.
Veröffentlicht: (2025)
von: Park, Junho, et al.
Veröffentlicht: (2025)
Object-Shot Enhanced Grounding Network for Egocentric Video
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
von: Feng, Yisen, et al.
Veröffentlicht: (2025)
EvObj: Learning Evolving Object-centric Representations for 3D Instance Segmentation without Scene Supervision
von: Chen, Jiahao, et al.
Veröffentlicht: (2026)
von: Chen, Jiahao, et al.
Veröffentlicht: (2026)
Benchmarking Egocentric Clinical Intent Understanding Capability for Medical Multimodal Large Language Models
von: Liu, Shaonan, et al.
Veröffentlicht: (2026)
von: Liu, Shaonan, et al.
Veröffentlicht: (2026)
Eyes on Target: Gaze-Aware Object Detection in Egocentric Video
von: Lall, Vishakha, et al.
Veröffentlicht: (2025)
von: Lall, Vishakha, et al.
Veröffentlicht: (2025)
ESP-PCT: Enhanced VR Semantic Performance through Efficient Compression of Temporal and Spatial Redundancies in Point Cloud Transformers
von: Mei, Luoyu, et al.
Veröffentlicht: (2024)
von: Mei, Luoyu, et al.
Veröffentlicht: (2024)
OSCaR: Object State Captioning and State Change Representation
von: Nguyen, Nguyen, et al.
Veröffentlicht: (2024)
von: Nguyen, Nguyen, et al.
Veröffentlicht: (2024)
EgoToM: Benchmarking Theory of Mind Reasoning from Egocentric Videos
von: Li, Yuxuan, et al.
Veröffentlicht: (2025)
von: Li, Yuxuan, et al.
Veröffentlicht: (2025)
Hierarchical Dual-Change Collaborative Learning for UAV Scene Change Captioning
von: Chen, Fuhai, et al.
Veröffentlicht: (2026)
von: Chen, Fuhai, et al.
Veröffentlicht: (2026)
Can We Change the Stroke Size for Easier Diffusion?
von: Bai, Yunwei, et al.
Veröffentlicht: (2026)
von: Bai, Yunwei, et al.
Veröffentlicht: (2026)
ObjectCompose: Evaluating Resilience of Vision-Based Models on Object-to-Background Compositional Changes
von: Malik, Hashmat Shadab, et al.
Veröffentlicht: (2024)
von: Malik, Hashmat Shadab, et al.
Veröffentlicht: (2024)
Changes in Gaza: DINOv3-Powered Multi-Class Change Detection for Damage Assessment in Conflict Zones
von: Zheng, Kai, et al.
Veröffentlicht: (2025)
von: Zheng, Kai, et al.
Veröffentlicht: (2025)
Leveraging Textual Compositional Reasoning for Robust Change Captioning
von: Park, Kyu Ri, et al.
Veröffentlicht: (2025)
von: Park, Kyu Ri, et al.
Veröffentlicht: (2025)
Predicting 3D Motion from 2D Video for Behavior-Based VR Biometrics
von: Li, Mingjun, et al.
Veröffentlicht: (2025)
von: Li, Mingjun, et al.
Veröffentlicht: (2025)
RoboSense: Large-scale Dataset and Benchmark for Egocentric Robot Perception and Navigation in Crowded and Unstructured Environments
von: Su, Haisheng, et al.
Veröffentlicht: (2024)
von: Su, Haisheng, et al.
Veröffentlicht: (2024)
Learning Egocentric In-Hand Object Segmentation through Weak Supervision from Human Narrations
von: Messina, Nicola, et al.
Veröffentlicht: (2025)
von: Messina, Nicola, et al.
Veröffentlicht: (2025)
TopV-Nav: Unlocking the Top-View Spatial Reasoning Potential of MLLM for Zero-shot Object Navigation
von: Zhong, Linqing, et al.
Veröffentlicht: (2024)
von: Zhong, Linqing, et al.
Veröffentlicht: (2024)
Towards user-centered interactive medical image segmentation in VR with an assistive AI agent
von: Spiegler, Pascal, et al.
Veröffentlicht: (2025)
von: Spiegler, Pascal, et al.
Veröffentlicht: (2025)
Ego-R1: Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning
von: Tian, Shulin, et al.
Veröffentlicht: (2025)
von: Tian, Shulin, et al.
Veröffentlicht: (2025)
Place-it-R1: Unlocking Environment-aware Reasoning Potential of MLLM for Video Object Insertion
von: Gu, Bohai, et al.
Veröffentlicht: (2026)
von: Gu, Bohai, et al.
Veröffentlicht: (2026)
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
von: Liu, Ye, et al.
Veröffentlicht: (2025)
von: Liu, Ye, et al.
Veröffentlicht: (2025)
Intention-Guided Cognitive Reasoning for Egocentric Long-Term Action Anticipation
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
von: Chu, Qiaohui, et al.
Veröffentlicht: (2025)
HAZARD Challenge: Embodied Decision Making in Dynamically Changing Environments
von: Zhou, Qinhong, et al.
Veröffentlicht: (2024)
von: Zhou, Qinhong, et al.
Veröffentlicht: (2024)
Cross-View Referring Multi-Object Tracking
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
von: Chen, Sijia, et al.
Veröffentlicht: (2024)
EgoEsportsQA: An Egocentric Video Benchmark for Perception and Reasoning in Esports
von: Ma, Jianzhe, et al.
Veröffentlicht: (2026)
von: Ma, Jianzhe, et al.
Veröffentlicht: (2026)
Improving Zero-Shot Object-Level Change Detection by Incorporating Visual Correspondence
von: Nguyen, Hung Huy, et al.
Veröffentlicht: (2025)
von: Nguyen, Hung Huy, et al.
Veröffentlicht: (2025)
ObjFiller3D: Scaling 3D Object Inpainting to Dense Multi-View Consistency
von: Feng, Haitang, et al.
Veröffentlicht: (2025)
von: Feng, Haitang, et al.
Veröffentlicht: (2025)
ChangeQuery: Advancing Remote Sensing Change Analysis for Natural and Human-Induced Disasters from Visual Detection to Semantic Understanding
von: Sun, Dongwei, et al.
Veröffentlicht: (2026)
von: Sun, Dongwei, et al.
Veröffentlicht: (2026)
A Mechanistic View on Video Generation as World Models: State and Dynamics
von: Wang, Luozhou, et al.
Veröffentlicht: (2026)
von: Wang, Luozhou, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
RAG-VR: Leveraging Retrieval-Augmented Generation for 3D Question Answering in VR Environments
von: Ding, Shiyi, et al.
Veröffentlicht: (2025) -
ObjFormer: Learning Land-Cover Changes From Paired OSM Data and Optical High-Resolution Imagery via Object-Guided Transformer
von: Chen, Hongruixuan, et al.
Veröffentlicht: (2023) -
MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning
von: Jiang, Zheng, et al.
Veröffentlicht: (2026) -
Fast Registration of Photorealistic Avatars for VR Facial Animation
von: Patel, Chaitanya, et al.
Veröffentlicht: (2024) -
OpenObj: Open-Vocabulary Object-Level Neural Radiance Fields with Fine-Grained Understanding
von: Deng, Yinan, et al.
Veröffentlicht: (2024)