Not All Pairs are Equal: Hierarchical Learning for Average-Precision-Oriented Video Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Yang, Xu, Qianqian, Wen, Peisong, Dai, Siran, Huang, Qingming |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When the Future Becomes the Past: Taming Temporal Correspondence for Self-supervised Video Representation Learning
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
Exploring Structural Degradation in Dense Representations for Self-supervised Learning
by: Dai, Siran, et al.
Published: (2025)
by: Dai, Siran, et al.
Published: (2025)
HiGFA: Hierarchical Guidance for Fine-grained Data Augmentation with Diffusion Models
by: Lu, Zhiguang, et al.
Published: (2025)
by: Lu, Zhiguang, et al.
Published: (2025)
Self-supervised Representation Learning with Local Aggregation for Image-based Profiling
by: Dai, Siran, et al.
Published: (2025)
by: Dai, Siran, et al.
Published: (2025)
From Static to Dynamic: Exploring Self-supervised Image-to-Video Representation Transfer Learning
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Semantic Concentration for Self-Supervised Dense Representations Learning
by: Wen, Peisong, et al.
Published: (2025)
by: Wen, Peisong, et al.
Published: (2025)
Bootstrapping Physics-Grounded Video Generation through VLM-Guided Iterative Self-Refinement
by: Liu, Yang, et al.
Published: (2025)
by: Liu, Yang, et al.
Published: (2025)
Regularized Contrastive Partial Multi-view Outlier Detection
by: Wang, Yijia, et al.
Published: (2024)
by: Wang, Yijia, et al.
Published: (2024)
Learn Faster and Remember More: Balancing Exploration and Exploitation for Continual Test-time Adaptation
by: Yang, Pinci, et al.
Published: (2025)
by: Yang, Pinci, et al.
Published: (2025)
Enhancing Sample Utilization in Noise-Robust Deep Metric Learning With Subgroup-Based Positive-Pair Selection
by: Yu, Zhipeng, et al.
Published: (2025)
by: Yu, Zhipeng, et al.
Published: (2025)
Top-K Pairwise Ranking: Bridging the Gap Among Ranking-Based Measures for Multi-Label Classification
by: Wang, Zitai, et al.
Published: (2024)
by: Wang, Zitai, et al.
Published: (2024)
AUCSeg: AUC-oriented Pixel-level Long-tail Semantic Segmentation
by: Han, Boyu, et al.
Published: (2024)
by: Han, Boyu, et al.
Published: (2024)
SOVC: Subject-Oriented Video Captioning
by: Teng, Chang, et al.
Published: (2023)
by: Teng, Chang, et al.
Published: (2023)
A Unified Perspective for Loss-Oriented Imbalanced Learning via Localization
by: Wang, Zitai, et al.
Published: (2023)
by: Wang, Zitai, et al.
Published: (2023)
Not All Diffusion Model Activations Have Been Evaluated as Discriminative Features
by: Meng, Benyuan, et al.
Published: (2024)
by: Meng, Benyuan, et al.
Published: (2024)
Bidirectional Logits Tree: Pursuing Granularity Reconcilement in Fine-Grained Classification
by: Lu, Zhiguang, et al.
Published: (2024)
by: Lu, Zhiguang, et al.
Published: (2024)
Not All Modalities Are Equal: Instruction-Aware Gating for Multimodal Videos
by: Ding, Bonan, et al.
Published: (2026)
by: Ding, Bonan, et al.
Published: (2026)
Not All Frame Features Are Equal: Video-to-4D Generation via Decoupling Dynamic-Static Features
by: Yang, Liying, et al.
Published: (2025)
by: Yang, Liying, et al.
Published: (2025)
DirMixE: Harnessing Test Agnostic Long-tail Recognition with Hierarchical Label Vartiations
by: Yang, Zhiyong, et al.
Published: (2024)
by: Yang, Zhiyong, et al.
Published: (2024)
Diffusion-based Adversarial Purification from the Perspective of the Frequency Domain
by: Pei, Gaozheng, et al.
Published: (2025)
by: Pei, Gaozheng, et al.
Published: (2025)
QVLA: Not All Channels Are Equal in Vision-Language-Action Model's Quantization
by: Xu, Yuhao, et al.
Published: (2026)
by: Xu, Yuhao, et al.
Published: (2026)
Making Training-Free Diffusion Segmentors Scale with the Generative Power
by: Meng, Benyuan, et al.
Published: (2026)
by: Meng, Benyuan, et al.
Published: (2026)
Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs
by: Xu, Zhikang, et al.
Published: (2026)
by: Xu, Zhikang, et al.
Published: (2026)
All Beings Are Equal in Open Set Recognition
by: Li, Chaohua, et al.
Published: (2024)
by: Li, Chaohua, et al.
Published: (2024)
RETTA: Retrieval-Enhanced Test-Time Adaptation for Zero-Shot Video Captioning
by: Ma, Yunchuan, et al.
Published: (2024)
by: Ma, Yunchuan, et al.
Published: (2024)
A Unified Framework for Stealthy Adversarial Generation via Latent Optimization and Transferability Enhancement
by: Pei, Gaozheng, et al.
Published: (2025)
by: Pei, Gaozheng, et al.
Published: (2025)
Size-invariance Matters: Rethinking Metrics and Losses for Imbalanced Multi-object Salient Object Detection
by: Li, Feiran, et al.
Published: (2024)
by: Li, Feiran, et al.
Published: (2024)
Uncertainty-boosted Robust Video Activity Anticipation
by: Qi, Zhaobo, et al.
Published: (2024)
by: Qi, Zhaobo, et al.
Published: (2024)
Not All Regions Are Equal: Attention-Guided Perturbation Network for Industrial Anomaly Detection
by: Huang, Tingfeng, et al.
Published: (2024)
by: Huang, Tingfeng, et al.
Published: (2024)
Not All Points Are Equal: Uncertainty-Aware 4D LiDAR Scene Synthesis
by: Xu, Xiang, et al.
Published: (2026)
by: Xu, Xiang, et al.
Published: (2026)
OphCLIP: Hierarchical Retrieval-Augmented Learning for Ophthalmic Surgical Video-Language Pretraining
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
Suppress Content Shift: Better Diffusion Features via Off-the-Shelf Generation Techniques
by: Meng, Benyuan, et al.
Published: (2024)
by: Meng, Benyuan, et al.
Published: (2024)
Not All Pixels Are Equal: Pixel-wise Meta-Learning for Medical Segmentation with Noisy Labels
by: Mu, Chenyu, et al.
Published: (2025)
by: Mu, Chenyu, et al.
Published: (2025)
Hybrid Generative Fusion for Efficient and Privacy-Preserving Face Recognition Dataset Generation
by: Li, Feiran, et al.
Published: (2025)
by: Li, Feiran, et al.
Published: (2025)
One Image is Worth a Thousand Words: A Usability Preservable Text-Image Collaborative Erasing Framework
by: Li, Feiran, et al.
Published: (2025)
by: Li, Feiran, et al.
Published: (2025)
Decorrelating Structure via Adapters Makes Ensemble Learning Practical for Semi-supervised Learning
by: Wu, Jiaqi, et al.
Published: (2024)
by: Wu, Jiaqi, et al.
Published: (2024)
Not All Pixels Are Equal: Confidence-Guided Attention for Feature Matching
by: Li, Dongyue
Published: (2025)
by: Li, Dongyue
Published: (2025)
HUD: Hierarchical Uncertainty-Aware Disambiguation Network for Composed Video Retrieval
by: Chen, Zhiwei, et al.
Published: (2025)
by: Chen, Zhiwei, et al.
Published: (2025)
Not All Tokens are Guided Equal: Improving Guidance in Visual Autoregressive Models
by: Nguyen, Ky Dan, et al.
Published: (2025)
by: Nguyen, Ky Dan, et al.
Published: (2025)
Local All-Pair Correspondence for Point Tracking
by: Cho, Seokju, et al.
Published: (2024)
by: Cho, Seokju, et al.
Published: (2024)
Similar Items
-
When the Future Becomes the Past: Taming Temporal Correspondence for Self-supervised Video Representation Learning
by: Liu, Yang, et al.
Published: (2025) -
Exploring Structural Degradation in Dense Representations for Self-supervised Learning
by: Dai, Siran, et al.
Published: (2025) -
HiGFA: Hierarchical Guidance for Fine-grained Data Augmentation with Diffusion Models
by: Lu, Zhiguang, et al.
Published: (2025) -
Self-supervised Representation Learning with Local Aggregation for Image-based Profiling
by: Dai, Siran, et al.
Published: (2025) -
From Static to Dynamic: Exploring Self-supervised Image-to-Video Representation Transfer Learning
by: Liu, Yang, et al.
Published: (2026)