Rethinking Unsupervised Cross-modal Flow Estimation: Learning from Decoupled Optimization and Consistency Constraint
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Runmin, Wang, Jialiang, Cao, Si-Yuan, Yu, Zhu, Yu, Junchen, Zhang, Guangyi, Shen, Hui-Liang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SSHNet: Unsupervised Cross-modal Homography Estimation via Problem Reformulation and Split Optimization
by: Yu, Junchen, et al.
Published: (2024)
by: Yu, Junchen, et al.
Published: (2024)
SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning
by: Zhang, Runmin, et al.
Published: (2024)
by: Zhang, Runmin, et al.
Published: (2024)
Boosting Multi-View Indoor 3D Object Detection via Adaptive 3D Volume Construction
by: Zhang, Runmin, et al.
Published: (2025)
by: Zhang, Runmin, et al.
Published: (2025)
Context and Geometry Aware Voxel Transformer for Semantic Scene Completion
by: Yu, Zhu, et al.
Published: (2024)
by: Yu, Zhu, et al.
Published: (2024)
EDFFDNet: Towards Accurate and Efficient Unsupervised Multi-Grid Image Registration
by: Zhu, Haokai, et al.
Published: (2025)
by: Zhu, Haokai, et al.
Published: (2025)
Rethinking Early-Fusion Strategies for Improved Multispectral Object Detection
by: Zhang, Xue, et al.
Published: (2024)
by: Zhang, Xue, et al.
Published: (2024)
Structure-Aware Radar-Camera Depth Estimation
by: Zhang, Fuyi, et al.
Published: (2025)
by: Zhang, Fuyi, et al.
Published: (2025)
SGDFormer: One-stage Transformer-based Architecture for Cross-Spectral Stereo Image Guided Denoising
by: Zhang, Runmin, et al.
Published: (2024)
by: Zhang, Runmin, et al.
Published: (2024)
Learned Image Transmission with Hierarchical Variational Autoencoder
by: Zhang, Guangyi, et al.
Published: (2024)
by: Zhang, Guangyi, et al.
Published: (2024)
Unsupervised Spike Depth Estimation via Cross-modality Cross-domain Knowledge Transfer
by: Liu, Jiaming, et al.
Published: (2022)
by: Liu, Jiaming, et al.
Published: (2022)
Large Depth Completion Model from Sparse Observations
by: Yu, Zhu, et al.
Published: (2026)
by: Yu, Zhu, et al.
Published: (2026)
Divide-and-Conquer Decoupled Network for Cross-Domain Few-Shot Segmentation
by: Cong, Runmin, et al.
Published: (2025)
by: Cong, Runmin, et al.
Published: (2025)
Language Driven Occupancy Prediction
by: Yu, Zhu, et al.
Published: (2024)
by: Yu, Zhu, et al.
Published: (2024)
Self-supervised Event-based Monocular Depth Estimation using Cross-modal Consistency
by: Zhu, Junyu, et al.
Published: (2024)
by: Zhu, Junyu, et al.
Published: (2024)
FlowVid: Taming Imperfect Optical Flows for Consistent Video-to-Video Synthesis
by: Liang, Feng, et al.
Published: (2023)
by: Liang, Feng, et al.
Published: (2023)
CNC: Cross-modal Normality Constraint for Unsupervised Multi-class Anomaly Detection
by: Wang, Xiaolei, et al.
Published: (2024)
by: Wang, Xiaolei, et al.
Published: (2024)
Decoupled Geometric Parameterization and its Application in Deep Homography Estimation
by: Huang, Yao, et al.
Published: (2025)
by: Huang, Yao, et al.
Published: (2025)
Improving Decoupled Posterior Sampling for Inverse Problems using Data Consistency Constraint
by: Qi, Zhi, et al.
Published: (2024)
by: Qi, Zhi, et al.
Published: (2024)
Deep Reversible Consistency Learning for Cross-modal Retrieval
by: Pu, Ruitao, et al.
Published: (2025)
by: Pu, Ruitao, et al.
Published: (2025)
Rethinking Cross-modal Interaction from a Top-down Perspective for Referring Video Object Segmentation
by: Liang, Chen, et al.
Published: (2021)
by: Liang, Chen, et al.
Published: (2021)
Boosting Instance Awareness via Cross-View Correlation with 4D Radar and Camera for 3D Object Detection
by: Bai, Xiaokai, et al.
Published: (2026)
by: Bai, Xiaokai, et al.
Published: (2026)
UTSRMorph: A Unified Transformer and Superresolution Network for Unsupervised Medical Image Registration
by: Zhang, Runshi, et al.
Published: (2024)
by: Zhang, Runshi, et al.
Published: (2024)
CDPR: Cross-modal Diffusion with Polarization for Reliable Monocular Depth Estimation
by: Yu, Rongjia, et al.
Published: (2026)
by: Yu, Rongjia, et al.
Published: (2026)
S-BEVLoc: BEV-based Self-supervised Framework for Large-scale LiDAR Global Localization
by: Zhang, Chenghao, et al.
Published: (2025)
by: Zhang, Chenghao, et al.
Published: (2025)
Learnable Cross-modal Knowledge Distillation for Multi-modal Learning with Missing Modality
by: Wang, Hu, et al.
Published: (2023)
by: Wang, Hu, et al.
Published: (2023)
Recurrent Cross-View Object Geo-Localization
by: Zhang, Xiaohan, et al.
Published: (2025)
by: Zhang, Xiaohan, et al.
Published: (2025)
From Contrast to Consistency: Rethinking Event-based Continuous-Time Optical Flow Estimation
by: Hu, Rui, et al.
Published: (2026)
by: Hu, Rui, et al.
Published: (2026)
Reasoning-Aligned Perception Decoupling for Scalable Multi-modal Reasoning
by: Gou, Yunhao, et al.
Published: (2025)
by: Gou, Yunhao, et al.
Published: (2025)
On Robust Cross-View Consistency in Self-Supervised Monocular Depth Estimation
by: Zhao, Haimei, et al.
Published: (2022)
by: Zhao, Haimei, et al.
Published: (2022)
Wan-Weaver: Interleaved Multi-modal Generation via Decoupled Training
by: Xing, Jinbo, et al.
Published: (2026)
by: Xing, Jinbo, et al.
Published: (2026)
Spectral Discrepancy and Cross-modal Semantic Consistency Learning for Object Detection in Hyperspectral Image
by: He, Xiao, et al.
Published: (2025)
by: He, Xiao, et al.
Published: (2025)
Cross-modal Prompting for Balanced Incomplete Multi-modal Emotion Recognition
by: He, Wen-Jue, et al.
Published: (2025)
by: He, Wen-Jue, et al.
Published: (2025)
Adversarially Masked Video Consistency for Unsupervised Domain Adaptation
by: Zhu, Xiaoyu, et al.
Published: (2024)
by: Zhu, Xiaoyu, et al.
Published: (2024)
RepVideo: Rethinking Cross-Layer Representation for Video Generation
by: Si, Chenyang, et al.
Published: (2025)
by: Si, Chenyang, et al.
Published: (2025)
Learning to Balance: Decoupled Siamese Diffusion Transformer for Reference-Based Remote Sensing Image Super-Resolution
by: Luo, Bin, et al.
Published: (2026)
by: Luo, Bin, et al.
Published: (2026)
Decoupled Spatio-Temporal Consistency Learning for Self-Supervised Tracking
by: Zheng, Yaozong, et al.
Published: (2025)
by: Zheng, Yaozong, et al.
Published: (2025)
CrossWeaver: Cross-modal Weaving for Arbitrary-Modality Semantic Segmentation
by: Zhang, Zelin, et al.
Published: (2026)
by: Zhang, Zelin, et al.
Published: (2026)
RestorerID: Towards Tuning-Free Face Restoration with ID Preservation
by: Ying, Jiacheng, et al.
Published: (2024)
by: Ying, Jiacheng, et al.
Published: (2024)
EventVGGT: Exploring Cross-Modal Distillation for Consistent Event-based Depth Estimation
by: Ren, Yinrui, et al.
Published: (2026)
by: Ren, Yinrui, et al.
Published: (2026)
Rethinking the Paradigm of Content Constraints in Unpaired Image-to-Image Translation
by: Cai, Xiuding, et al.
Published: (2022)
by: Cai, Xiuding, et al.
Published: (2022)
Similar Items
-
SSHNet: Unsupervised Cross-modal Homography Estimation via Problem Reformulation and Split Optimization
by: Yu, Junchen, et al.
Published: (2024) -
SCPNet: Unsupervised Cross-modal Homography Estimation via Intra-modal Self-supervised Learning
by: Zhang, Runmin, et al.
Published: (2024) -
Boosting Multi-View Indoor 3D Object Detection via Adaptive 3D Volume Construction
by: Zhang, Runmin, et al.
Published: (2025) -
Context and Geometry Aware Voxel Transformer for Semantic Scene Completion
by: Yu, Zhu, et al.
Published: (2024) -
EDFFDNet: Towards Accurate and Efficient Unsupervised Multi-Grid Image Registration
by: Zhu, Haokai, et al.
Published: (2025)