SceneDiff: A Benchmark and Method for Multiview Object Change Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Yuqun, Lin, Chih-hao, Che, Henry, Tiwari, Aditi, Zou, Chuhang, Wang, Shenlong, Hoiem, Derek |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MonoPatchNeRF: Improving Neural Radiance Fields with Patch-based Monocular Guidance
by: Wu, Yuqun, et al.
Published: (2024)
by: Wu, Yuqun, et al.
Published: (2024)
Plenoptic PNG: Real-Time Neural Radiance Fields in 150 KB
by: Lee, Jae Yong, et al.
Published: (2024)
by: Lee, Jae Yong, et al.
Published: (2024)
TextRegion: Text-Aligned Region Tokens from Frozen Image-Text Models
by: Xiao, Yao, et al.
Published: (2025)
by: Xiao, Yao, et al.
Published: (2025)
Multiview Scene Graph
by: Zhang, Juexiao, et al.
Published: (2024)
by: Zhang, Juexiao, et al.
Published: (2024)
LidarDM: Generative LiDAR Simulation in a Generated World
by: Zyrianov, Vlas, et al.
Published: (2024)
by: Zyrianov, Vlas, et al.
Published: (2024)
Anytime Continual Learning for Open Vocabulary Classification
by: Zhu, Zhen, et al.
Published: (2024)
by: Zhu, Zhen, et al.
Published: (2024)
Toward Realistic Camouflaged Object Detection: Benchmarks and Method
by: Xin, Zhimeng, et al.
Published: (2025)
by: Xin, Zhimeng, et al.
Published: (2025)
RELOCATE: A Simple Training-Free Baseline for Visual Query Localization Using Region-Based Representations
by: Khosla, Savya, et al.
Published: (2024)
by: Khosla, Savya, et al.
Published: (2024)
Continual Learning in Open-vocabulary Classification with Complementary Memory Systems
by: Zhu, Zhen, et al.
Published: (2023)
by: Zhu, Zhen, et al.
Published: (2023)
ACT360: An Efficient 360-Degree Action Detection and Summarization Framework for Mission-Critical Training and Debriefing
by: Tiwari, Aditi, et al.
Published: (2025)
by: Tiwari, Aditi, et al.
Published: (2025)
Region-Based Representations Revisited
by: Shlapentokh-Rothman, Michal, et al.
Published: (2024)
by: Shlapentokh-Rothman, Michal, et al.
Published: (2024)
InvRGB+L: Inverse Rendering of Complex Scenes with Unified Color and LiDAR Reflectance Modeling
by: Chen, Xiaoxue, et al.
Published: (2025)
by: Chen, Xiaoxue, et al.
Published: (2025)
Stepper: Stepwise Immersive Scene Generation with Multiview Panoramas
by: Wimbauer, Felix, et al.
Published: (2026)
by: Wimbauer, Felix, et al.
Published: (2026)
Geometry-Aware Diffusion Models for Multiview Scene Inpainting
by: Salimi, Ahmad, et al.
Published: (2025)
by: Salimi, Ahmad, et al.
Published: (2025)
Visual Program Distillation with Template-Based Augmentation
by: Shlapentokh-Rothman, Michal, et al.
Published: (2024)
by: Shlapentokh-Rothman, Michal, et al.
Published: (2024)
Select-Mosaic: Data Augmentation Method for Dense Small Object Scenes
by: Zhang, Hao, et al.
Published: (2024)
by: Zhang, Hao, et al.
Published: (2024)
HoloScene: Simulation-Ready Interactive 3D Worlds from a Single Video
by: Xia, Hongchi, et al.
Published: (2025)
by: Xia, Hongchi, et al.
Published: (2025)
Evaluating the Performance of Open-Vocabulary Object Detection in Low-quality Image
by: Wu, Po-Chih
Published: (2025)
by: Wu, Po-Chih
Published: (2025)
T-REN: Learning Text-Aligned Region Tokens Improves Dense Vision-Language Alignment and Scalability
by: Khosla, Savya, et al.
Published: (2026)
by: Khosla, Savya, et al.
Published: (2026)
REN: Fast and Efficient Region Encodings from Patch-Based Image Encoders
by: Khosla, Savya, et al.
Published: (2025)
by: Khosla, Savya, et al.
Published: (2025)
Evaluating Multiview Object Consistency in Humans and Image Models
by: Bonnen, Tyler, et al.
Published: (2024)
by: Bonnen, Tyler, et al.
Published: (2024)
UrbanIR: Large-Scale Urban Scene Inverse Rendering from a Single Video
by: Lin, Chih-Hao, et al.
Published: (2023)
by: Lin, Chih-Hao, et al.
Published: (2023)
Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval
by: Shlapentokh-Rothman, Michal, et al.
Published: (2026)
by: Shlapentokh-Rothman, Michal, et al.
Published: (2026)
Generalizable Sparse-View 3D Reconstruction from Unconstrained Images
by: Gupta, Vinayak, et al.
Published: (2026)
by: Gupta, Vinayak, et al.
Published: (2026)
IRIS: Inverse Rendering of Indoor Scenes from Low Dynamic Range Images
by: Lin, Chih-Hao, et al.
Published: (2024)
by: Lin, Chih-Hao, et al.
Published: (2024)
Credible Teacher for Semi-Supervised Object Detection in Open Scene
by: Zhuang, Jingyu, et al.
Published: (2024)
by: Zhuang, Jingyu, et al.
Published: (2024)
Benchmarking Single-Step Inpainting Methods for Multi-Object 3D Gaussian Splatting Scenes
by: Dröge, Finn, et al.
Published: (2026)
by: Dröge, Finn, et al.
Published: (2026)
Controllable Video Object Insertion via Multiview Priors
by: Qi, Xia, et al.
Published: (2026)
by: Qi, Xia, et al.
Published: (2026)
Language-driven Description Generation and Common Sense Reasoning for Video Action Recognition
by: Hu, Xiaodan, et al.
Published: (2025)
by: Hu, Xiaodan, et al.
Published: (2025)
Towards Reflected Object Detection: A Benchmark
by: Wu, Yiquan, et al.
Published: (2024)
by: Wu, Yiquan, et al.
Published: (2024)
Scene Adaptive Sparse Transformer for Event-based Object Detection
by: Peng, Yansong, et al.
Published: (2024)
by: Peng, Yansong, et al.
Published: (2024)
DCHM: Depth-Consistent Human Modeling for Multiview Detection
by: Ma, Jiahao, et al.
Published: (2025)
by: Ma, Jiahao, et al.
Published: (2025)
Sparse Multiview Open-Vocabulary 3D Detection
by: Moliner, Olivier, et al.
Published: (2025)
by: Moliner, Olivier, et al.
Published: (2025)
SceneEdited: A City-Scale Benchmark for 3D HD Map Updating via Image-Guided Change Detection
by: Lin, Chun-Jung, et al.
Published: (2025)
by: Lin, Chun-Jung, et al.
Published: (2025)
SiM3D: Single-instance Multiview Multimodal and Multisetup 3D Anomaly Detection Benchmark
by: Costanzino, Alex, et al.
Published: (2025)
by: Costanzino, Alex, et al.
Published: (2025)
How to Teach Large Multimodal Models New Skills
by: Zhu, Zhen, et al.
Published: (2025)
by: Zhu, Zhen, et al.
Published: (2025)
FoBa: A Foreground-Background co-Guided Method and New Benchmark for Remote Sensing Semantic Change Detection
by: Zhang, Haotian, et al.
Published: (2025)
by: Zhang, Haotian, et al.
Published: (2025)
Zero-Shot Scene Change Detection
by: Cho, Kyusik, et al.
Published: (2024)
by: Cho, Kyusik, et al.
Published: (2024)
Language-Driven Object-Oriented Two-Stage Method for Scene Graph Anticipation
by: Zhu, Xiaomeng, et al.
Published: (2025)
by: Zhu, Xiaomeng, et al.
Published: (2025)
Object Style Diffusion for Generalized Object Detection in Urban Scene
by: Li, Hao, et al.
Published: (2024)
by: Li, Hao, et al.
Published: (2024)
Similar Items
-
MonoPatchNeRF: Improving Neural Radiance Fields with Patch-based Monocular Guidance
by: Wu, Yuqun, et al.
Published: (2024) -
Plenoptic PNG: Real-Time Neural Radiance Fields in 150 KB
by: Lee, Jae Yong, et al.
Published: (2024) -
TextRegion: Text-Aligned Region Tokens from Frozen Image-Text Models
by: Xiao, Yao, et al.
Published: (2025) -
Multiview Scene Graph
by: Zhang, Juexiao, et al.
Published: (2024) -
LidarDM: Generative LiDAR Simulation in a Generated World
by: Zyrianov, Vlas, et al.
Published: (2024)