Multi-granularity Correspondence Learning from Long-term Noisy Videos
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lin, Yijie, Zhang, Jie, Huang, Zhenyu, Liu, Jia, Wen, Zujie, Peng, Xi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Noisy-Correspondence Learning for Text-to-Image Person Re-identification
von: Qin, Yang, et al.
Veröffentlicht: (2023)
von: Qin, Yang, et al.
Veröffentlicht: (2023)
Dual-granularity Sinkhorn Distillation for Enhanced Learning from Long-tailed Noisy Data
von: Hong, Feng, et al.
Veröffentlicht: (2025)
von: Hong, Feng, et al.
Veröffentlicht: (2025)
Disentangled Noisy Correspondence Learning
von: Dang, Zhuohang, et al.
Veröffentlicht: (2024)
von: Dang, Zhuohang, et al.
Veröffentlicht: (2024)
Multi-granularity Contrastive Cross-modal Collaborative Generation for End-to-End Long-term Video Question Answering
von: Yu, Ting, et al.
Veröffentlicht: (2024)
von: Yu, Ting, et al.
Veröffentlicht: (2024)
SALI: Short-term Alignment and Long-term Interaction Network for Colonoscopy Video Polyp Segmentation
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
von: Hu, Qiang, et al.
Veröffentlicht: (2024)
PCSR: Pseudo-label Consistency-Guided Sample Refinement for Noisy Correspondence Learning
von: Liu, Zhuoyao, et al.
Veröffentlicht: (2025)
von: Liu, Zhuoyao, et al.
Veröffentlicht: (2025)
Mitigating Noisy Correspondence by Geometrical Structure Consistency Learning
von: Zhao, Zihua, et al.
Veröffentlicht: (2024)
von: Zhao, Zihua, et al.
Veröffentlicht: (2024)
Robust Noisy Correspondence Learning via Self-Drop and Dual-Weight
von: Liu, Fan, et al.
Veröffentlicht: (2024)
von: Liu, Fan, et al.
Veröffentlicht: (2024)
ReCon: Enhancing True Correspondence Discrimination through Relation Consistency for Robust Noisy Correspondence Learning
von: Zha, Quanxing, et al.
Veröffentlicht: (2025)
von: Zha, Quanxing, et al.
Veröffentlicht: (2025)
When the Future Becomes the Past: Taming Temporal Correspondence for Self-supervised Video Representation Learning
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
Multi-Sentence Grounding for Long-term Instructional Video
von: Li, Zeqian, et al.
Veröffentlicht: (2023)
von: Li, Zeqian, et al.
Veröffentlicht: (2023)
Decoupled Contrastive Multi-View Clustering with High-Order Random Walks
von: Lu, Yiding, et al.
Veröffentlicht: (2023)
von: Lu, Yiding, et al.
Veröffentlicht: (2023)
Learning to Generate Diverse Pedestrian Movements from Web Videos with Noisy Labels
von: Liu, Zhizheng, et al.
Veröffentlicht: (2024)
von: Liu, Zhizheng, et al.
Veröffentlicht: (2024)
LLaVA-ReID: Selective Multi-image Questioner for Interactive Person Re-Identification
von: Lu, Yiding, et al.
Veröffentlicht: (2025)
von: Lu, Yiding, et al.
Veröffentlicht: (2025)
Dynamic Uncertainty Learning with Noisy Correspondence for Text-Based Person Search
von: Xie, Zequn, et al.
Veröffentlicht: (2025)
von: Xie, Zequn, et al.
Veröffentlicht: (2025)
REPAIR: Rank Correlation and Noisy Pair Half-replacing with Memory for Noisy Correspondence
von: Zheng, Ruochen, et al.
Veröffentlicht: (2024)
von: Zheng, Ruochen, et al.
Veröffentlicht: (2024)
Semantics Meets Temporal Correspondence: Self-supervised Object-centric Learning in Videos
von: Qian, Rui, et al.
Veröffentlicht: (2023)
von: Qian, Rui, et al.
Veröffentlicht: (2023)
Video World Models with Long-term Spatial Memory
von: Wu, Tong, et al.
Veröffentlicht: (2025)
von: Wu, Tong, et al.
Veröffentlicht: (2025)
Hallucination Mitigation Prompts Long-term Video Understanding
von: Sun, Yiwei, et al.
Veröffentlicht: (2024)
von: Sun, Yiwei, et al.
Veröffentlicht: (2024)
MUVR: A Multi-Modal Untrimmed Video Retrieval Benchmark with Multi-Level Visual Correspondence
von: Feng, Yue, et al.
Veröffentlicht: (2025)
von: Feng, Yue, et al.
Veröffentlicht: (2025)
Unlearning the Noisy Correspondence Makes CLIP More Robust
von: Han, Haochen, et al.
Veröffentlicht: (2025)
von: Han, Haochen, et al.
Veröffentlicht: (2025)
Backdooring Self-Supervised Contrastive Learning by Noisy Alignment
von: Chen, Tuo, et al.
Veröffentlicht: (2025)
von: Chen, Tuo, et al.
Veröffentlicht: (2025)
PicoPose: Progressive Pixel-to-Pixel Correspondence Learning for Novel Object Pose Estimation
von: Liu, Lihua, et al.
Veröffentlicht: (2025)
von: Liu, Lihua, et al.
Veröffentlicht: (2025)
Knowledge Distillation with Multi-granularity Mixture of Priors for Image Super-Resolution
von: Li, Simiao, et al.
Veröffentlicht: (2024)
von: Li, Simiao, et al.
Veröffentlicht: (2024)
Robust Remote Sensing Image-Text Retrieval with Noisy Correspondence
von: Song, Qiya, et al.
Veröffentlicht: (2026)
von: Song, Qiya, et al.
Veröffentlicht: (2026)
Mavors: Multi-granularity Video Representation for Multimodal Large Language Model
von: Shi, Yang, et al.
Veröffentlicht: (2025)
von: Shi, Yang, et al.
Veröffentlicht: (2025)
X-ReID: Multi-granularity Information Interaction for Video-Based Visible-Infrared Person Re-Identification
von: Yu, Chenyang, et al.
Veröffentlicht: (2025)
von: Yu, Chenyang, et al.
Veröffentlicht: (2025)
State-space Decomposition Model for Video Prediction Considering Long-term Motion Trend
von: Cui, Fei, et al.
Veröffentlicht: (2024)
von: Cui, Fei, et al.
Veröffentlicht: (2024)
CARE: Class-Adaptive Expert Consensus for Reliable Learning with Long-Tailed Noisy Labels
von: Li, Mengke, et al.
Veröffentlicht: (2026)
von: Li, Mengke, et al.
Veröffentlicht: (2026)
Holmes-VAU: Towards Long-term Video Anomaly Understanding at Any Granularity
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
von: Zhang, Huaxin, et al.
Veröffentlicht: (2024)
Cross-modal Active Complementary Learning with Self-refining Correspondence
von: Qin, Yang, et al.
Veröffentlicht: (2023)
von: Qin, Yang, et al.
Veröffentlicht: (2023)
Attention-Driven Multimodal Alignment for Long-term Action Quality Assessment
von: Wang, Xin, et al.
Veröffentlicht: (2025)
von: Wang, Xin, et al.
Veröffentlicht: (2025)
Robust Ego-Exo Correspondence with Long-Term Memory
von: Hu, Yijun, et al.
Veröffentlicht: (2025)
von: Hu, Yijun, et al.
Veröffentlicht: (2025)
Zero-Shot Video Translation and Editing with Frame Spatial-Temporal Correspondence
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
Towards Effective Long-Video Event Prediction via Multi-Level Event Semantics Mining
von: Peng, Bo, et al.
Veröffentlicht: (2026)
von: Peng, Bo, et al.
Veröffentlicht: (2026)
Fusion of Short-term and Long-term Attention for Video Mirror Detection
von: Xu, Mingchen, et al.
Veröffentlicht: (2024)
von: Xu, Mingchen, et al.
Veröffentlicht: (2024)
FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting
von: He, Zefeng, et al.
Veröffentlicht: (2025)
von: He, Zefeng, et al.
Veröffentlicht: (2025)
MiLA: Multi-view Intensive-fidelity Long-term Video Generation World Model for Autonomous Driving
von: Wang, Haiguang, et al.
Veröffentlicht: (2025)
von: Wang, Haiguang, et al.
Veröffentlicht: (2025)
Learning Long-form Video Prior via Generative Pre-Training
von: Xie, Jinheng, et al.
Veröffentlicht: (2024)
von: Xie, Jinheng, et al.
Veröffentlicht: (2024)
LongVPO: From Anchored Cues to Self-Reasoning for Long-Form Video Preference Optimization
von: Huang, Zhenpeng, et al.
Veröffentlicht: (2026)
von: Huang, Zhenpeng, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Noisy-Correspondence Learning for Text-to-Image Person Re-identification
von: Qin, Yang, et al.
Veröffentlicht: (2023) -
Dual-granularity Sinkhorn Distillation for Enhanced Learning from Long-tailed Noisy Data
von: Hong, Feng, et al.
Veröffentlicht: (2025) -
Disentangled Noisy Correspondence Learning
von: Dang, Zhuohang, et al.
Veröffentlicht: (2024) -
Multi-granularity Contrastive Cross-modal Collaborative Generation for End-to-End Long-term Video Question Answering
von: Yu, Ting, et al.
Veröffentlicht: (2024) -
SALI: Short-term Alignment and Long-term Interaction Network for Colonoscopy Video Polyp Segmentation
von: Hu, Qiang, et al.
Veröffentlicht: (2024)