Towards Visual Discrimination and Reasoning of Real-World Physical Dynamics: Physics-Grounded Anomaly Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Wenqiao, Gu, Yao, Chen, Xintao, Xu, Xiaohao, Hu, Ming, Huang, Xiaonan, Wu, Yingna |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multi-turn Physics-informed Vision-language Model for Physics-grounded Anomaly Detection
von: Gu, Yao, et al.
Veröffentlicht: (2026)
von: Gu, Yao, et al.
Veröffentlicht: (2026)
Bridging 3D Anomaly Localization and Repair via High-Quality Continuous Geometric Representation
von: Zheng, Bozhong, et al.
Veröffentlicht: (2025)
von: Zheng, Bozhong, et al.
Veröffentlicht: (2025)
Unsupervised Multi-View Visual Anomaly Detection via Progressive Homography-Guided Alignment
von: Chen, Xintao, et al.
Veröffentlicht: (2025)
von: Chen, Xintao, et al.
Veröffentlicht: (2025)
Customizing Visual-Language Foundation Models for Multi-modal Anomaly Detection and Reasoning
von: Xu, Xiaohao, et al.
Veröffentlicht: (2024)
von: Xu, Xiaohao, et al.
Veröffentlicht: (2024)
Multi-Sensor Object Anomaly Detection: Unifying Appearance, Geometry, and Internal Properties
von: Li, Wenqiao, et al.
Veröffentlicht: (2024)
von: Li, Wenqiao, et al.
Veröffentlicht: (2024)
Photorealistic Phantom Roads in Real Scenes: Disentangling 3D Hallucinations from Physical Geometry
von: Nguyen, Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Hoang, et al.
Veröffentlicht: (2025)
Breaking the Rigid Prior: Towards Articulated 3D Anomaly Detection
von: Gan, Jinye, et al.
Veröffentlicht: (2026)
von: Gan, Jinye, et al.
Veröffentlicht: (2026)
A Survey on Visual Anomaly Detection: Challenge, Approach, and Prospect
von: Cao, Yunkang, et al.
Veröffentlicht: (2024)
von: Cao, Yunkang, et al.
Veröffentlicht: (2024)
Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration
von: Wang, Jun, et al.
Veröffentlicht: (2026)
von: Wang, Jun, et al.
Veröffentlicht: (2026)
Visual Anomaly Detection under Complex View-Illumination Interplay: A Large-Scale Benchmark
von: Cao, Yunkang, et al.
Veröffentlicht: (2025)
von: Cao, Yunkang, et al.
Veröffentlicht: (2025)
Holmes-VAD: Towards Unbiased and Explainable Video Anomaly Detection via Multi-modal LLM
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
Holmes-VAU: Towards Long-term Video Anomaly Understanding at Any Granularity
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
GlanceVAD: Exploring Glance Supervision for Label-efficient Video Anomaly Detection
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing
von: Xu, Ruihang, et al.
Veröffentlicht: (2026)
von: Xu, Ruihang, et al.
Veröffentlicht: (2026)
Semantic Visual Anomaly Detection and Reasoning in AI-Generated Images
von: Tan, Chuangchuang, et al.
Veröffentlicht: (2025)
von: Tan, Chuangchuang, et al.
Veröffentlicht: (2025)
Is Your LiDAR Placement Optimized for 3D Scene Understanding?
von: Li, Ye, et al.
Veröffentlicht: (2024)
von: Li, Ye, et al.
Veröffentlicht: (2024)
AnomalyMoE: Towards a Language-free Generalist Model for Unified Visual Anomaly Detection
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2025)
von: Gu, Zhaopeng, et al.
Veröffentlicht: (2025)
MAC-Ego3D: Multi-Agent Gaussian Consensus for Real-Time Collaborative Ego-Motion and Photorealistic 3D Reconstruction
von: Xu, Xiaohao, et al.
Veröffentlicht: (2024)
von: Xu, Xiaohao, et al.
Veröffentlicht: (2024)
Complementary Pseudo Multimodal Feature for Point Cloud Anomaly Detection
von: Cao, Yunkang, et al.
Veröffentlicht: (2023)
von: Cao, Yunkang, et al.
Veröffentlicht: (2023)
GeoWeaver: Grounding Visual Tokens with Geometric Evidence before Scene Reasoning
von: Miao, Deshui, et al.
Veröffentlicht: (2026)
von: Miao, Deshui, et al.
Veröffentlicht: (2026)
Learn Suspected Anomalies from Event Prompts for Video Anomaly Detection
von: Tao, Chenchen, et al.
Veröffentlicht: (2024)
von: Tao, Chenchen, et al.
Veröffentlicht: (2024)
LogiCode: an LLM-Driven Framework for Logical Anomaly Detection
von: Zhang, Yiheng, et al.
Veröffentlicht: (2024)
von: Zhang, Yiheng, et al.
Veröffentlicht: (2024)
DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World
von: Li, Xiangtai, et al.
Veröffentlicht: (2025)
von: Li, Xiangtai, et al.
Veröffentlicht: (2025)
DMAD: Dual Memory Bank for Real-World Anomaly Detection
von: Hu, Jianlong, et al.
Veröffentlicht: (2024)
von: Hu, Jianlong, et al.
Veröffentlicht: (2024)
When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills
von: Wang, Yunfei, et al.
Veröffentlicht: (2026)
von: Wang, Yunfei, et al.
Veröffentlicht: (2026)
Real-IAD: A Real-World Multi-View Dataset for Benchmarking Versatile Industrial Anomaly Detection
von: Wang, Chengjie, et al.
Veröffentlicht: (2024)
von: Wang, Chengjie, et al.
Veröffentlicht: (2024)
OpenGround: Active Cognition-based Reasoning for Open-World 3D Visual Grounding
von: Huang, Wenyuan, et al.
Veröffentlicht: (2025)
von: Huang, Wenyuan, et al.
Veröffentlicht: (2025)
MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning
von: Wu, Chengfei, et al.
Veröffentlicht: (2025)
von: Wu, Chengfei, et al.
Veröffentlicht: (2025)
CLIP-Flow: A Universal Discriminator for AI-Generated Images Inspired by Anomaly Detection
von: Yuan, Zhipeng, et al.
Veröffentlicht: (2025)
von: Yuan, Zhipeng, et al.
Veröffentlicht: (2025)
PhyGround: Benchmarking Physical Reasoning in Generative World Models
von: Lin, Juyi, et al.
Veröffentlicht: (2026)
von: Lin, Juyi, et al.
Veröffentlicht: (2026)
EgoCoT-Bench: Benchmarking Grounded and Verifiable Operation-Centric Chain of Thought Reasoning for MLLMs
von: Dai, Yang, et al.
Veröffentlicht: (2026)
von: Dai, Yang, et al.
Veröffentlicht: (2026)
RGBT-Ground Benchmark: Visual Grounding Beyond RGB in Complex Real-World Scenarios
von: Zhao, Tianyi, et al.
Veröffentlicht: (2025)
von: Zhao, Tianyi, et al.
Veröffentlicht: (2025)
Sparse Reasoning is Enough: Biological-Inspired Framework for Video Anomaly Detection with Large Pre-trained Models
von: Huang, He, et al.
Veröffentlicht: (2025)
von: Huang, He, et al.
Veröffentlicht: (2025)
BARE: Towards Bias-Aware and Reasoning-Enhanced One-Tower Visual Grounding
von: Li, Hongbing, et al.
Veröffentlicht: (2026)
von: Li, Hongbing, et al.
Veröffentlicht: (2026)
Thinking in Dynamics: How Multimodal Large Language Models Perceive, Track, and Reason Dynamics in Physical 4D World
von: Huang, Yuzhi, et al.
Veröffentlicht: (2026)
von: Huang, Yuzhi, et al.
Veröffentlicht: (2026)
Grounding Video Reasoning in Physical Signals
von: Osmanli, Alibay, et al.
Veröffentlicht: (2026)
von: Osmanli, Alibay, et al.
Veröffentlicht: (2026)
VisualTrans: A Benchmark for Real-World Visual Transformation Reasoning
von: Ji, Yuheng, et al.
Veröffentlicht: (2025)
von: Ji, Yuheng, et al.
Veröffentlicht: (2025)
A Recover-then-Discriminate Framework for Robust Anomaly Detection
von: Xing, Peng, et al.
Veröffentlicht: (2024)
von: Xing, Peng, et al.
Veröffentlicht: (2024)
Towards Transferable Targeted 3D Adversarial Attack in the Physical World
von: Huang, Yao, et al.
Veröffentlicht: (2023)
von: Huang, Yao, et al.
Veröffentlicht: (2023)
PhysicsMind: Sim and Real Mechanics Benchmarking for Physical Reasoning and Prediction in Foundational VLMs and World Models
von: Mak, Chak-Wing, et al.
Veröffentlicht: (2026)
von: Mak, Chak-Wing, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Multi-turn Physics-informed Vision-language Model for Physics-grounded Anomaly Detection
von: Gu, Yao, et al.
Veröffentlicht: (2026) -
Bridging 3D Anomaly Localization and Repair via High-Quality Continuous Geometric Representation
von: Zheng, Bozhong, et al.
Veröffentlicht: (2025) -
Unsupervised Multi-View Visual Anomaly Detection via Progressive Homography-Guided Alignment
von: Chen, Xintao, et al.
Veröffentlicht: (2025) -
Customizing Visual-Language Foundation Models for Multi-modal Anomaly Detection and Reasoning
von: Xu, Xiaohao, et al.
Veröffentlicht: (2024) -
Multi-Sensor Object Anomaly Detection: Unifying Appearance, Geometry, and Internal Properties
von: Li, Wenqiao, et al.
Veröffentlicht: (2024)