Learning Cross-View Object Correspondence via Cycle-Consistent Mask Prediction
Fuente:
arXiv
Salvato in:
| Autori principali: | Yan, Shannan, Zheng, Leqi, Lv, Keyu, Ni, Jingchen, Wei, Hongyang, Zhang, Jiajun, Wang, Guangting, Lyu, Jing, Yuan, Chun, Rao, Fengyun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AdaMem: Adaptive User-Centric Memory for Long-Horizon Dialogue Agents
di: Yan, Shannan, et al.
Pubblicazione: (2026)
di: Yan, Shannan, et al.
Pubblicazione: (2026)
What Makes Low-Bit Quantization-Aware Training Work for Reasoning LLMs? A Systematic Study
di: Lv, Keyu, et al.
Pubblicazione: (2026)
di: Lv, Keyu, et al.
Pubblicazione: (2026)
FCL-COD: Weakly Supervised Camouflaged Object Detection with Frequency-aware and Contrastive Learning
di: Ni, Jingchen, et al.
Pubblicazione: (2026)
di: Ni, Jingchen, et al.
Pubblicazione: (2026)
ObjEmbed: Towards Universal Multimodal Object Embeddings
di: Fu, Shenghao, et al.
Pubblicazione: (2026)
di: Fu, Shenghao, et al.
Pubblicazione: (2026)
WeDetect: Fast Open-Vocabulary Object Detection as Retrieval
di: Fu, Shenghao, et al.
Pubblicazione: (2025)
di: Fu, Shenghao, et al.
Pubblicazione: (2025)
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots
di: Zheng, Guangting, et al.
Pubblicazione: (2025)
di: Zheng, Guangting, et al.
Pubblicazione: (2025)
HQ-CLIP: Leveraging Large Vision-Language Models to Create High-Quality Image-Text Datasets and CLIP Models
di: Wei, Zhixiang, et al.
Pubblicazione: (2025)
di: Wei, Zhixiang, et al.
Pubblicazione: (2025)
Mask Consistency Regularization in Object Removal
di: Yuan, Hua, et al.
Pubblicazione: (2025)
di: Yuan, Hua, et al.
Pubblicazione: (2025)
D-ORCA: Dialogue-Centric Optimization for Robust Audio-Visual Captioning
di: Tang, Changli, et al.
Pubblicazione: (2026)
di: Tang, Changli, et al.
Pubblicazione: (2026)
S3R-GS: Streamlining the Pipeline for Large-Scale Street Scene Reconstruction
di: Zheng, Guangting, et al.
Pubblicazione: (2025)
di: Zheng, Guangting, et al.
Pubblicazione: (2025)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
di: Huang, Kuan Wei, et al.
Pubblicazione: (2025)
di: Huang, Kuan Wei, et al.
Pubblicazione: (2025)
Visual Perception by Large Language Model's Weights
di: Ma, Feipeng, et al.
Pubblicazione: (2024)
di: Ma, Feipeng, et al.
Pubblicazione: (2024)
Multi-Modal Generative Embedding Model
di: Ma, Feipeng, et al.
Pubblicazione: (2024)
di: Ma, Feipeng, et al.
Pubblicazione: (2024)
Rethinking Temporal Consistency in Video Object-Centric Learning: From Prediction to Correspondence
di: Li, Zhiyuan, et al.
Pubblicazione: (2026)
di: Li, Zhiyuan, et al.
Pubblicazione: (2026)
Public attitudes toward personalized medicine in breast cancer
di: Shannan, Ghassan
Pubblicazione: (2026)
di: Shannan, Ghassan
Pubblicazione: (2026)
El sonido y la furia: la persuasión multicultural en México y Estados Unidos, por José Antonio Aguilar Rivera, México, Santillana Ediciones Generales, 2004, 292 p.
di: Shannan Mattiace
Pubblicazione: (2007)
di: Shannan Mattiace
Pubblicazione: (2007)
Reformas multiculturales para los mayas de Yucatán
di: Shannan Mattiace
Pubblicazione: (2015)
di: Shannan Mattiace
Pubblicazione: (2015)
MMhops-R1: Multimodal Multi-hop Reasoning
di: Zhang, Tao, et al.
Pubblicazione: (2025)
di: Zhang, Tao, et al.
Pubblicazione: (2025)
On Robust Cross-View Consistency in Self-Supervised Monocular Depth Estimation
di: Zhao, Haimei, et al.
Pubblicazione: (2022)
di: Zhao, Haimei, et al.
Pubblicazione: (2022)
A Dynamical Framework for the McKay Correspondence via Gauge-Theoretic Morse Flow
di: Yan, Jiajun
Pubblicazione: (2026)
di: Yan, Jiajun
Pubblicazione: (2026)
IteRPrimE: Zero-shot Referring Image Segmentation with Iterative Grad-CAM Refinement and Primary Word Emphasis
di: Wang, Yuji, et al.
Pubblicazione: (2025)
di: Wang, Yuji, et al.
Pubblicazione: (2025)
3D-Consistent Multi-View Editing by Correspondence Guidance
di: Bengtson, Josef, et al.
Pubblicazione: (2025)
di: Bengtson, Josef, et al.
Pubblicazione: (2025)
Towards Cross-View Point Correspondence in Vision-Language Models
di: Wang, Yipu, et al.
Pubblicazione: (2025)
di: Wang, Yipu, et al.
Pubblicazione: (2025)
What Should I Cite? A RAG Benchmark for Academic Citation Prediction
di: Zheng, Leqi, et al.
Pubblicazione: (2026)
di: Zheng, Leqi, et al.
Pubblicazione: (2026)
WonderFree: Enhancing Novel View Quality and Cross-View Consistency for 3D Scene Exploration
di: Ni, Chaojun, et al.
Pubblicazione: (2025)
di: Ni, Chaojun, et al.
Pubblicazione: (2025)
V$^{2}$-SAM: Marrying SAM2 with Multi-Prompt Experts for Cross-View Object Correspondence
di: Pan, Jiancheng, et al.
Pubblicazione: (2025)
di: Pan, Jiancheng, et al.
Pubblicazione: (2025)
CycleBEV: Regularizing View Transformation Networks via View Cycle Consistency for Bird's-Eye-View Semantic Segmentation
di: Hong, Jeongbin, et al.
Pubblicazione: (2026)
di: Hong, Jeongbin, et al.
Pubblicazione: (2026)
InfoGeo: Information-Theoretic Object-Centric Learning for Cross-View Generalizable UAV Geo-Localization
di: Zhang, Hongyang, et al.
Pubblicazione: (2026)
di: Zhang, Hongyang, et al.
Pubblicazione: (2026)
Cycle Consistency in Video Object-Centric Learning
di: Zhao, Rongzhen, et al.
Pubblicazione: (2026)
di: Zhao, Rongzhen, et al.
Pubblicazione: (2026)
VAGeo: View-specific Attention for Cross-View Object Geo-Localization
di: Li, Zhongyang, et al.
Pubblicazione: (2025)
di: Li, Zhongyang, et al.
Pubblicazione: (2025)
REVERSE: Reinforcing Evidence Verification and Search for Agentic Image geo-localization
di: Li, Yong, et al.
Pubblicazione: (2026)
di: Li, Yong, et al.
Pubblicazione: (2026)
Cross-View Completion Models are Zero-shot Correspondence Estimators
di: An, Honggyu, et al.
Pubblicazione: (2024)
di: An, Honggyu, et al.
Pubblicazione: (2024)
MOGeo: Beyond One-to-One Cross-View Object Geo-localization
di: Lv, Bo, et al.
Pubblicazione: (2026)
di: Lv, Bo, et al.
Pubblicazione: (2026)
The Split Janus-faced Sun: Magnetic Rhythm and Duality in the Solar Cycle
di: Chen, Weiqi, et al.
Pubblicazione: (2026)
di: Chen, Weiqi, et al.
Pubblicazione: (2026)
Self-Supervised Partial Cycle-Consistency for Multi-View Matching
di: Taggenbrock, Fedor, et al.
Pubblicazione: (2025)
di: Taggenbrock, Fedor, et al.
Pubblicazione: (2025)
Advancements in Monomolecular Multimodal Platforms for Cancer Theranostics
di: Shannan, Ghassan, et al.
Pubblicazione: (2025)
di: Shannan, Ghassan, et al.
Pubblicazione: (2025)
Do LLMs Overcome Shortcut Learning? An Evaluation of Shortcut Challenges in Large Language Models
di: Yuan, Yu, et al.
Pubblicazione: (2024)
di: Yuan, Yu, et al.
Pubblicazione: (2024)
CorrespondentDream: Enhancing 3D Fidelity of Text-to-3D using Cross-View Correspondences
di: Kim, Seungwook, et al.
Pubblicazione: (2024)
di: Kim, Seungwook, et al.
Pubblicazione: (2024)
Recurrent Cross-View Object Geo-Localization
di: Zhang, Xiaohan, et al.
Pubblicazione: (2025)
di: Zhang, Xiaohan, et al.
Pubblicazione: (2025)
MDE-Edit: Masked Dual-Editing for Multi-Object Image Editing via Diffusion Models
di: Zhu, Hongyang, et al.
Pubblicazione: (2025)
di: Zhu, Hongyang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
AdaMem: Adaptive User-Centric Memory for Long-Horizon Dialogue Agents
di: Yan, Shannan, et al.
Pubblicazione: (2026) -
What Makes Low-Bit Quantization-Aware Training Work for Reasoning LLMs? A Systematic Study
di: Lv, Keyu, et al.
Pubblicazione: (2026) -
FCL-COD: Weakly Supervised Camouflaged Object Detection with Frequency-aware and Contrastive Learning
di: Ni, Jingchen, et al.
Pubblicazione: (2026) -
ObjEmbed: Towards Universal Multimodal Object Embeddings
di: Fu, Shenghao, et al.
Pubblicazione: (2026) -
WeDetect: Fast Open-Vocabulary Object Detection as Retrieval
di: Fu, Shenghao, et al.
Pubblicazione: (2025)