Learning Compatible Multi-Prize Subnetworks for Asymmetric Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Yushuai, Zhou, Zikun, Jiang, Dongmei, Wang, Yaowei, Yu, Jun, Lu, Guangming, Pei, Wenjie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Prototype Perturbation for Relaxing Alignment Constraints in Backward-Compatible Learning
von: Zhou, Zikun, et al.
Veröffentlicht: (2025)
von: Zhou, Zikun, et al.
Veröffentlicht: (2025)
Beyond Heuristics: Learnable Density Control for 3D Gaussian Splatting
von: Ning, Zhenhua, et al.
Veröffentlicht: (2026)
von: Ning, Zhenhua, et al.
Veröffentlicht: (2026)
EditInfinity: Image Editing with Binary-Quantized Generative Models
von: Wang, Jiahuan, et al.
Veröffentlicht: (2025)
von: Wang, Jiahuan, et al.
Veröffentlicht: (2025)
D$^2$ST-Adapter: Disentangled-and-Deformable Spatio-Temporal Adapter for Few-shot Action Recognition
von: Pei, Wenjie, et al.
Veröffentlicht: (2023)
von: Pei, Wenjie, et al.
Veröffentlicht: (2023)
Video-ToC: Video Tree-of-Cue Reasoning
von: Tan, Qizhong, et al.
Veröffentlicht: (2026)
von: Tan, Qizhong, et al.
Veröffentlicht: (2026)
Prune Redundancy, Preserve Essence: Vision Token Compression in VLMs via Synergistic Importance-Diversity
von: Fang, Zhengyao, et al.
Veröffentlicht: (2026)
von: Fang, Zhengyao, et al.
Veröffentlicht: (2026)
Deeply-Conditioned Image Compression via Self-Generated Priors
von: Zhao, Zhineng, et al.
Veröffentlicht: (2025)
von: Zhao, Zhineng, et al.
Veröffentlicht: (2025)
DiffTrans: Differentiable Geometry-Materials Decomposition for Reconstructing Transparent Objects
von: Li, Changpu, et al.
Veröffentlicht: (2026)
von: Li, Changpu, et al.
Veröffentlicht: (2026)
CricaVPR: Cross-image Correlation-aware Representation Learning for Visual Place Recognition
von: Lu, Feng, et al.
Veröffentlicht: (2024)
von: Lu, Feng, et al.
Veröffentlicht: (2024)
Harnessing Vision-Language Pretrained Models with Temporal-Aware Adaptation for Referring Video Object Segmentation
von: Zhou, Zikun, et al.
Veröffentlicht: (2024)
von: Zhou, Zikun, et al.
Veröffentlicht: (2024)
Recognition-Synergistic Scene Text Editing
von: Fang, Zhengyao, et al.
Veröffentlicht: (2025)
von: Fang, Zhengyao, et al.
Veröffentlicht: (2025)
Learning to Rebalance Multi-Modal Optimization by Adaptively Masking Subnetworks
von: Yang, Yang, et al.
Veröffentlicht: (2024)
von: Yang, Yang, et al.
Veröffentlicht: (2024)
DS-Det: Single-Query Paradigm and Attention Disentangled Learning for Flexible Object Detection
von: Cao, Guiping, et al.
Veröffentlicht: (2025)
von: Cao, Guiping, et al.
Veröffentlicht: (2025)
Global-Local Stepwise Generative Network for Ultra High-Resolution Image Restoration
von: Feng, Xin, et al.
Veröffentlicht: (2022)
von: Feng, Xin, et al.
Veröffentlicht: (2022)
Domain-Rectifying Adapter for Cross-Domain Few-Shot Segmentation
von: Su, Jiapeng, et al.
Veröffentlicht: (2024)
von: Su, Jiapeng, et al.
Veröffentlicht: (2024)
RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation
von: Wen, Junwei, et al.
Veröffentlicht: (2026)
von: Wen, Junwei, et al.
Veröffentlicht: (2026)
UniVoxel: Fast Inverse Rendering by Unified Voxelization of Scene Representation
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
von: Wu, Shuang, et al.
Veröffentlicht: (2024)
MixBCT: Towards Self-Adapting Backward-Compatible Training
von: Liang, Yu, et al.
Veröffentlicht: (2023)
von: Liang, Yu, et al.
Veröffentlicht: (2023)
RTracker: Recoverable Tracking via PN Tree Structured Memory
von: Huang, Yuqing, et al.
Veröffentlicht: (2024)
von: Huang, Yuqing, et al.
Veröffentlicht: (2024)
Too Vivid to Be Real? Benchmarking and Calibrating Generative Color Fidelity
von: Fang, Zhengyao, et al.
Veröffentlicht: (2026)
von: Fang, Zhengyao, et al.
Veröffentlicht: (2026)
Prompt Customization for Continual Learning
von: Dai, Yong, et al.
Veröffentlicht: (2024)
von: Dai, Yong, et al.
Veröffentlicht: (2024)
Enhancing Spatial Reasoning in Multimodal Large Language Models through Reasoning-based Segmentation
von: Ning, Zhenhua, et al.
Veröffentlicht: (2025)
von: Ning, Zhenhua, et al.
Veröffentlicht: (2025)
Cross-DINO: Cross the Deep MLP and Transformer for Small Object Detection
von: Cao, Guiping, et al.
Veröffentlicht: (2025)
von: Cao, Guiping, et al.
Veröffentlicht: (2025)
Instruction-guided Multi-Granularity Segmentation and Captioning with Large Multimodal Model
von: Zhou, Li, et al.
Veröffentlicht: (2024)
von: Zhou, Li, et al.
Veröffentlicht: (2024)
EMMA: Empowering Multi-modal Mamba with Structural and Hierarchical Alignment
von: Xing, Yifei, et al.
Veröffentlicht: (2024)
von: Xing, Yifei, et al.
Veröffentlicht: (2024)
Efficient Adversarial Training via Criticality-Aware Fine-Tuning
von: Li, Wenyun, et al.
Veröffentlicht: (2026)
von: Li, Wenyun, et al.
Veröffentlicht: (2026)
Motion-aware Latent Diffusion Models for Video Frame Interpolation
von: Huang, Zhilin, et al.
Veröffentlicht: (2024)
von: Huang, Zhilin, et al.
Veröffentlicht: (2024)
Multi-scale Contrastive Adaptor Learning for Segmenting Anything in Underperformed Scenes
von: Zhou, Ke, et al.
Veröffentlicht: (2024)
von: Zhou, Ke, et al.
Veröffentlicht: (2024)
WeCromCL: Weakly Supervised Cross-Modality Contrastive Learning for Transcription-only Supervised Text Spotting
von: Wu, Jingjing, et al.
Veröffentlicht: (2024)
von: Wu, Jingjing, et al.
Veröffentlicht: (2024)
SOLAR: Self-supervised Joint Learning for Symmetric Multimodal Retrieval
von: Yang, Wenjie, et al.
Veröffentlicht: (2026)
von: Yang, Wenjie, et al.
Veröffentlicht: (2026)
CLIP-based Synergistic Knowledge Transfer for Text-based Person Retrieval
von: Liu, Yating, et al.
Veröffentlicht: (2023)
von: Liu, Yating, et al.
Veröffentlicht: (2023)
UP-Person: Unified Parameter-Efficient Transfer Learning for Text-based Person Retrieval
von: Liu, Yating, et al.
Veröffentlicht: (2025)
von: Liu, Yating, et al.
Veröffentlicht: (2025)
Image2Sentence based Asymmetrical Zero-shot Composed Image Retrieval
von: Du, Yongchao, et al.
Veröffentlicht: (2024)
von: Du, Yongchao, et al.
Veröffentlicht: (2024)
RMLer: Synthesizing Novel Objects across Diverse Categories via Reinforcement Mixing Learning
von: Li, Jun, et al.
Veröffentlicht: (2025)
von: Li, Jun, et al.
Veröffentlicht: (2025)
RadioDUN: A Physics-Inspired Deep Unfolding Network for Radio Map Estimation
von: Chen, Taiqin, et al.
Veröffentlicht: (2025)
von: Chen, Taiqin, et al.
Veröffentlicht: (2025)
Codebook-Based Adaptive Feature Compression With Semantic Enhancement for Edge-Cloud Systems
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
von: Wang, Xinyu, et al.
Veröffentlicht: (2025)
AHDMIL: Asymmetric Hierarchical Distillation Multi-Instance Learning for Fast and Accurate Whole-Slide Image Classification
von: Dong, Jiuyang, et al.
Veröffentlicht: (2025)
von: Dong, Jiuyang, et al.
Veröffentlicht: (2025)
MergeOcc: Bridge the Domain Gap between Different LiDARs for Robust Occupancy Prediction
von: Xu, Zikun, et al.
Veröffentlicht: (2024)
von: Xu, Zikun, et al.
Veröffentlicht: (2024)
HLFormer: Enhancing Partially Relevant Video Retrieval with Hyperbolic Learning
von: Li, Jun, et al.
Veröffentlicht: (2025)
von: Li, Jun, et al.
Veröffentlicht: (2025)
Prompt-Driven Dynamic Object-Centric Learning for Single Domain Generalization
von: Li, Deng, et al.
Veröffentlicht: (2024)
von: Li, Deng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Prototype Perturbation for Relaxing Alignment Constraints in Backward-Compatible Learning
von: Zhou, Zikun, et al.
Veröffentlicht: (2025) -
Beyond Heuristics: Learnable Density Control for 3D Gaussian Splatting
von: Ning, Zhenhua, et al.
Veröffentlicht: (2026) -
EditInfinity: Image Editing with Binary-Quantized Generative Models
von: Wang, Jiahuan, et al.
Veröffentlicht: (2025) -
D$^2$ST-Adapter: Disentangled-and-Deformable Spatio-Temporal Adapter for Few-shot Action Recognition
von: Pei, Wenjie, et al.
Veröffentlicht: (2023) -
Video-ToC: Video Tree-of-Cue Reasoning
von: Tan, Qizhong, et al.
Veröffentlicht: (2026)