ME-CPT: Multi-Task Enhanced Cross-Temporal Point Transformer for Urban 3D Change Detection
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Luqi, Wang, Haiping, Liu, Chong, Dong, Zhen, Yang, Bisheng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DPG-CD: Depth-Prior-Guided Cross-Modal Joint 2D-3D Change Detection
by: Zhang, Luqi, et al.
Published: (2026)
by: Zhang, Luqi, et al.
Published: (2026)
SpatialLLM: From Multi-modality Data to Urban Spatial Intelligence
by: Chen, Jiabin, et al.
Published: (2025)
by: Chen, Jiabin, et al.
Published: (2025)
SVII-3D: Advancing Roadside Infrastructure Inventory with Decimeter-level 3D Localization and Comprehension from Sparse Street Imagery
by: Liu, Chong, et al.
Published: (2026)
by: Liu, Chong, et al.
Published: (2026)
FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators
by: Wang, Haiping, et al.
Published: (2023)
by: Wang, Haiping, et al.
Published: (2023)
Unleashing the Capabilities of Large Vision-Language Models for Intelligent Perception of Roadside Infrastructure
by: Fu, Luxuan, et al.
Published: (2026)
by: Fu, Luxuan, et al.
Published: (2026)
Explicitly Guided Information Interaction Network for Cross-modal Point Cloud Completion
by: Xu, Hang, et al.
Published: (2024)
by: Xu, Hang, et al.
Published: (2024)
VistaDream: Sampling multiview consistent images for single-view scene reconstruction
by: Wang, Haiping, et al.
Published: (2024)
by: Wang, Haiping, et al.
Published: (2024)
GAGS: Granularity-Aware Feature Distillation for Language Gaussian Splatting
by: Peng, Yuning, et al.
Published: (2024)
by: Peng, Yuning, et al.
Published: (2024)
Exploiting Motion Prior for Accurate Pose Estimation of Dashboard Cameras
by: Lu, Yipeng, et al.
Published: (2024)
by: Lu, Yipeng, et al.
Published: (2024)
PTT: Point-Trajectory Transformer for Efficient Temporal 3D Object Detection
by: Huang, Kuan-Chih, et al.
Published: (2023)
by: Huang, Kuan-Chih, et al.
Published: (2023)
Continuous Urban Change Detection from Satellite Image Time Series with Temporal Feature Refinement and Multi-Task Integration
by: Hafner, Sebastian, et al.
Published: (2024)
by: Hafner, Sebastian, et al.
Published: (2024)
Temporal-Anchor3DLane: Enhanced 3D Lane Detection with Multi-Task Losses and LSTM Fusion
by: Suhas, D. Shainu, et al.
Published: (2025)
by: Suhas, D. Shainu, et al.
Published: (2025)
CoreNet: Conflict Resolution Network for Point-Pixel Misalignment and Sub-Task Suppression of 3D LiDAR-Camera Object Detection
by: Li, Yiheng, et al.
Published: (2025)
by: Li, Yiheng, et al.
Published: (2025)
SPPSFormer: High-quality Superpoint-based Transformer for Roof Plane Instance Segmentation from Point Clouds
by: Zeng, Cheng, et al.
Published: (2025)
by: Zeng, Cheng, et al.
Published: (2025)
LifelongPR: Lifelong point cloud place recognition based on sample replay and prompt learning
by: Zou, Xianghong, et al.
Published: (2025)
by: Zou, Xianghong, et al.
Published: (2025)
HyPCV-Former: Hyperbolic Spatio-Temporal Transformer for 3D Point Cloud Video Anomaly Detection
by: Cao, Jiaping, et al.
Published: (2025)
by: Cao, Jiaping, et al.
Published: (2025)
WHU-STree: A Multi-modal Benchmark Dataset for Street Tree Inventory
by: Ding, Ruifei, et al.
Published: (2025)
by: Ding, Ruifei, et al.
Published: (2025)
Expert Knowledge-Guided Decision Calibration for Accurate Fine-Grained Tree Species Classification
by: Long, Chen, et al.
Published: (2026)
by: Long, Chen, et al.
Published: (2026)
Mobile-Seed: Joint Semantic Segmentation and Boundary Detection for Mobile Robots
by: Liao, Youqi, et al.
Published: (2023)
by: Liao, Youqi, et al.
Published: (2023)
OU-CoViT: Copula-Enhanced Bi-Channel Multi-Task Vision Transformers with Dual Adaptation for OU-UWF Images
by: Li, Yang, et al.
Published: (2024)
by: Li, Yang, et al.
Published: (2024)
ChangeNet: Multi-Temporal Asymmetric Change Detection Dataset
by: Ji, Deyi, et al.
Published: (2023)
by: Ji, Deyi, et al.
Published: (2023)
DeepAAT: Deep Automated Aerial Triangulation for Fast UAV-based Mapping
by: Chen, Zequan, et al.
Published: (2024)
by: Chen, Zequan, et al.
Published: (2024)
Radar-Camera BEV Multi-Task Learning with Cross-Task Attention Bridge for Joint 3D Detection and Segmentation
by: İnanç, Ahmet, et al.
Published: (2026)
by: İnanç, Ahmet, et al.
Published: (2026)
Temporal-Aware Spiking Transformer Hashing Based on 3D-DWT
by: Mei, Zihao, et al.
Published: (2025)
by: Mei, Zihao, et al.
Published: (2025)
PVAFN: Point-Voxel Attention Fusion Network with Multi-Pooling Enhancing for 3D Object Detection
by: Li, Yidi, et al.
Published: (2024)
by: Li, Yidi, et al.
Published: (2024)
V2X-R: Cooperative LiDAR-4D Radar Fusion with Denoising Diffusion for 3D Object Detection
by: Huang, Xun, et al.
Published: (2024)
by: Huang, Xun, et al.
Published: (2024)
DAOcc: 3D Object Detection Assisted Multi-Sensor Fusion for 3D Occupancy Prediction
by: Yang, Zhen, et al.
Published: (2024)
by: Yang, Zhen, et al.
Published: (2024)
Enhanced Parking Perception by Multi-Task Fisheye Cross-view Transformers
by: Musabini, Antonyo, et al.
Published: (2024)
by: Musabini, Antonyo, et al.
Published: (2024)
Spatial-Temporal Graph Enhanced DETR Towards Multi-Frame 3D Object Detection
by: Zhang, Yifan, et al.
Published: (2023)
by: Zhang, Yifan, et al.
Published: (2023)
Reliable-loc: Robust sequential LiDAR global localization in large-scale street scenes based on verifiable cues
by: Zou, Xianghong, et al.
Published: (2024)
by: Zou, Xianghong, et al.
Published: (2024)
SaliencyI2PLoc: saliency-guided image-point cloud localization using contrastive learning
by: Li, Yuhao, et al.
Published: (2024)
by: Li, Yuhao, et al.
Published: (2024)
PEP-GS: Perceptually-Enhanced Precise Structured 3D Gaussians for View-Adaptive Rendering
by: Jin, Junxi, et al.
Published: (2024)
by: Jin, Junxi, et al.
Published: (2024)
STeInFormer: Spatial-Temporal Interaction Transformer Architecture for Remote Sensing Change Detection
by: Ma, Xiaowen, et al.
Published: (2024)
by: Ma, Xiaowen, et al.
Published: (2024)
PVTransformer: Point-to-Voxel Transformer for Scalable 3D Object Detection
by: Leng, Zhaoqi, et al.
Published: (2024)
by: Leng, Zhaoqi, et al.
Published: (2024)
StereoMV2D: A Sparse Temporal Stereo-Enhanced Framework for Robust Multi-View 3D Object Detection
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
3D-RAD: A Comprehensive 3D Radiology Med-VQA Dataset with Multi-Temporal Analysis and Diverse Diagnostic Tasks
by: Gai, Xiaotang, et al.
Published: (2025)
by: Gai, Xiaotang, et al.
Published: (2025)
CPT-Interp: Continuous sPatial and Temporal Motion Modeling for 4D Medical Image Interpolation
by: Li, Xia, et al.
Published: (2024)
by: Li, Xia, et al.
Published: (2024)
OSMLoc: Single Image-Based Visual Localization in OpenStreetMap with Fused Geometric and Semantic Guidance
by: Liao, Youqi, et al.
Published: (2024)
by: Liao, Youqi, et al.
Published: (2024)
Distilling Temporal Knowledge with Masked Feature Reconstruction for 3D Object Detection
by: Zheng, Haowen, et al.
Published: (2024)
by: Zheng, Haowen, et al.
Published: (2024)
Temporal-Enhanced Multimodal Transformer for Referring Multi-Object Tracking and Segmentation
by: Xiao, Changcheng, et al.
Published: (2024)
by: Xiao, Changcheng, et al.
Published: (2024)
Similar Items
-
DPG-CD: Depth-Prior-Guided Cross-Modal Joint 2D-3D Change Detection
by: Zhang, Luqi, et al.
Published: (2026) -
SpatialLLM: From Multi-modality Data to Urban Spatial Intelligence
by: Chen, Jiabin, et al.
Published: (2025) -
SVII-3D: Advancing Roadside Infrastructure Inventory with Decimeter-level 3D Localization and Comprehension from Sparse Street Imagery
by: Liu, Chong, et al.
Published: (2026) -
FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators
by: Wang, Haiping, et al.
Published: (2023) -
Unleashing the Capabilities of Large Vision-Language Models for Intelligent Perception of Roadside Infrastructure
by: Fu, Luxuan, et al.
Published: (2026)