Synergistic Perception and Generative Recomposition: A Multi-Agent Orchestration for Expert-Level Building Inspection
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhong, Hui, Gao, Yichun, Liu, Luyan, Guo, Xusen, Kuang, Zhaonian, Zhang, Qiming, Zheng, Xinhu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
di: Kuang, Zhaonian, et al.
Pubblicazione: (2026)
di: Kuang, Zhaonian, et al.
Pubblicazione: (2026)
Can Large Multimodal Models Inspect Buildings? A Hierarchical Benchmark for Structural Pathology Reasoning
di: Zhong, Hui, et al.
Pubblicazione: (2026)
di: Zhong, Hui, et al.
Pubblicazione: (2026)
Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
di: Ding, Rui, et al.
Pubblicazione: (2026)
di: Ding, Rui, et al.
Pubblicazione: (2026)
RayD3D: Distilling Depth Knowledge Along the Ray for Robust Multi-View 3D Object Detection
di: Ding, Rui, et al.
Pubblicazione: (2026)
di: Ding, Rui, et al.
Pubblicazione: (2026)
CoIn3D: Revisiting Configuration-Invariant Multi-Camera 3D Object Detection
di: Kuang, Zhaonian, et al.
Pubblicazione: (2026)
di: Kuang, Zhaonian, et al.
Pubblicazione: (2026)
UniT: Unified Geometry Learning with Group Autoregressive Transformer
di: Wang, Haotian, et al.
Pubblicazione: (2026)
di: Wang, Haotian, et al.
Pubblicazione: (2026)
Scope: Selective Cross-modal Orchestration of Visual Perception Experts
di: Zhang, Tianyu, et al.
Pubblicazione: (2025)
di: Zhang, Tianyu, et al.
Pubblicazione: (2025)
Multi-Modal Building Inspection via Perceiver IO Fusion of Satellite and Street-Level Imagery
di: Sombekke, Niels, et al.
Pubblicazione: (2026)
di: Sombekke, Niels, et al.
Pubblicazione: (2026)
FaceBench: A Multi-View Multi-Level Facial Attribute VQA Dataset for Benchmarking Face Perception MLLMs
di: Wang, Xiaoqin, et al.
Pubblicazione: (2025)
di: Wang, Xiaoqin, et al.
Pubblicazione: (2025)
ConMo: Controllable Motion Disentanglement and Recomposition for Zero-Shot Motion Transfer
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
MARBLE: Material Recomposition and Blending in CLIP-Space
di: Cheng, Ta-Ying, et al.
Pubblicazione: (2025)
di: Cheng, Ta-Ying, et al.
Pubblicazione: (2025)
Hollywood Town: Long-Video Generation via Cross-Modal Multi-Agent Orchestration
di: Wei, Zheng, et al.
Pubblicazione: (2025)
di: Wei, Zheng, et al.
Pubblicazione: (2025)
Modular Embedding Recomposition for Incremental Learning
di: Panariello, Aniello, et al.
Pubblicazione: (2025)
di: Panariello, Aniello, et al.
Pubblicazione: (2025)
Triamese-ViT: A 3D-Aware Method for Robust Brain Age Estimation from MRIs
di: Zhang, Zhaonian, et al.
Pubblicazione: (2024)
di: Zhang, Zhaonian, et al.
Pubblicazione: (2024)
MCE: Towards a General Framework for Handling Missing Modalities under Imbalanced Missing Rates
di: Zhao, Binyu, et al.
Pubblicazione: (2025)
di: Zhao, Binyu, et al.
Pubblicazione: (2025)
Frequency-Domain Decomposition and Recomposition for Robust Audio-Visual Segmentation
di: Shen, Yunzhe, et al.
Pubblicazione: (2025)
di: Shen, Yunzhe, et al.
Pubblicazione: (2025)
Generalized Fine-Grained Category Discovery with Multi-Granularity Conceptual Experts
di: Zheng, Haiyang, et al.
Pubblicazione: (2025)
di: Zheng, Haiyang, et al.
Pubblicazione: (2025)
Multi-Level Feature Fusion for Continual Learning in Visual Quality Inspection
di: Bauer, Johannes C., et al.
Pubblicazione: (2026)
di: Bauer, Johannes C., et al.
Pubblicazione: (2026)
A New Clustering-based View Planning Method for Building Inspection with Drone
di: Zheng, Yongshuai, et al.
Pubblicazione: (2024)
di: Zheng, Yongshuai, et al.
Pubblicazione: (2024)
CAMEO: A Conditional and Quality-Aware Multi-Agent Image Editing Orchestrator
di: Pu, Yuhan, et al.
Pubblicazione: (2026)
di: Pu, Yuhan, et al.
Pubblicazione: (2026)
InstanceRSR: Real-World Super-Resolution via Instance-Aware Representation Alignment
di: Guo, Zixin, et al.
Pubblicazione: (2026)
di: Guo, Zixin, et al.
Pubblicazione: (2026)
BOOKAGENT: Orchestrating Safety-Aware Visual Narratives via Multi-Agent Cognitive Calibration
di: Gao, Bo, et al.
Pubblicazione: (2026)
di: Gao, Bo, et al.
Pubblicazione: (2026)
ORXE: Orchestrating Experts for Dynamically Configurable Efficiency
di: Wang, Qingyuan, et al.
Pubblicazione: (2025)
di: Wang, Qingyuan, et al.
Pubblicazione: (2025)
Exploring Fine-Grained Representation and Recomposition for Cloth-Changing Person Re-Identification
di: Wang, Qizao, et al.
Pubblicazione: (2023)
di: Wang, Qizao, et al.
Pubblicazione: (2023)
CoLC: Communication-Efficient Collaborative Perception with LiDAR Completion
di: Han, Yushan, et al.
Pubblicazione: (2026)
di: Han, Yushan, et al.
Pubblicazione: (2026)
SoftHGNN: Soft Hypergraph Neural Networks for General Visual Recognition
di: Lei, Mengqi, et al.
Pubblicazione: (2025)
di: Lei, Mengqi, et al.
Pubblicazione: (2025)
ShowUI-Aloha: Human-Taught GUI Agent
di: Zhang, Yichun, et al.
Pubblicazione: (2026)
di: Zhang, Yichun, et al.
Pubblicazione: (2026)
YOLOv13: Real-Time Object Detection with Hypergraph-Enhanced Adaptive Visual Perception
di: Lei, Mengqi, et al.
Pubblicazione: (2025)
di: Lei, Mengqi, et al.
Pubblicazione: (2025)
On the Global Photometric Alignment for Low-Level Vision
di: Li, Mingjia, et al.
Pubblicazione: (2026)
di: Li, Mingjia, et al.
Pubblicazione: (2026)
BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections
di: Varghese, Subin, et al.
Pubblicazione: (2025)
di: Varghese, Subin, et al.
Pubblicazione: (2025)
Scale Propagation Network for Generalizable Depth Completion
di: Wang, Haotian, et al.
Pubblicazione: (2024)
di: Wang, Haotian, et al.
Pubblicazione: (2024)
Orchestrate Latent Expertise: Advancing Online Continual Learning with Multi-Level Supervision and Reverse Self-Distillation
di: Yan, HongWei, et al.
Pubblicazione: (2024)
di: Yan, HongWei, et al.
Pubblicazione: (2024)
Multi-view Image Prompted Multi-view Diffusion for Improved 3D Generation
di: Kim, Seungwook, et al.
Pubblicazione: (2024)
di: Kim, Seungwook, et al.
Pubblicazione: (2024)
EyePCR: A Comprehensive Benchmark for Fine-Grained Perception, Knowledge Comprehension and Clinical Reasoning in Ophthalmic Surgery
di: Wang, Gui, et al.
Pubblicazione: (2025)
di: Wang, Gui, et al.
Pubblicazione: (2025)
AgentAlign: Misalignment-Adapted Multi-Agent Perception for Resilient Inter-Agent Sensor Correlations
di: Meng, Zonglin, et al.
Pubblicazione: (2024)
di: Meng, Zonglin, et al.
Pubblicazione: (2024)
GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation
di: Chen, Sixiang, et al.
Pubblicazione: (2026)
di: Chen, Sixiang, et al.
Pubblicazione: (2026)
AdvGPS: Adversarial GPS for Multi-Agent Perception Attack
di: Li, Jinlong, et al.
Pubblicazione: (2024)
di: Li, Jinlong, et al.
Pubblicazione: (2024)
ROVI: A VLM-LLM Re-Captioned Dataset for Open-Vocabulary Instance-Grounded Text-to-Image Generation
di: Peng, Cihang, et al.
Pubblicazione: (2025)
di: Peng, Cihang, et al.
Pubblicazione: (2025)
CC-FMO: Camera-Conditioned Zero-Shot Single Image to 3D Scene Generation with Foundation Model Orchestration
di: Tang, Boshi, et al.
Pubblicazione: (2025)
di: Tang, Boshi, et al.
Pubblicazione: (2025)
CoReS: Orchestrating the Dance of Reasoning and Segmentation
di: Bao, Xiaoyi, et al.
Pubblicazione: (2024)
di: Bao, Xiaoyi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
di: Kuang, Zhaonian, et al.
Pubblicazione: (2026) -
Can Large Multimodal Models Inspect Buildings? A Hierarchical Benchmark for Structural Pathology Reasoning
di: Zhong, Hui, et al.
Pubblicazione: (2026) -
Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
di: Ding, Rui, et al.
Pubblicazione: (2026) -
RayD3D: Distilling Depth Knowledge Along the Ray for Robust Multi-View 3D Object Detection
di: Ding, Rui, et al.
Pubblicazione: (2026) -
CoIn3D: Revisiting Configuration-Invariant Multi-Camera 3D Object Detection
di: Kuang, Zhaonian, et al.
Pubblicazione: (2026)