Decoupled Multi-Predictor Optimization for Inference-Efficient Model Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Luo, Liwei, Li, Shuaitengyuan, Ren, Dongwei, Wang, Qilong, Zhu, Pengfei, Hu, Qinghua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AMU-Tuning: Effective Logit Bias for CLIP-based Few-shot Learning
by: Tang, Yuwei, et al.
Published: (2024)
by: Tang, Yuwei, et al.
Published: (2024)
Unleashing Degradation-Carrying Features in Symmetric U-Net: Simpler and Stronger Baselines for All-in-One Image Restoration
by: Jiao, Wenlong, et al.
Published: (2025)
by: Jiao, Wenlong, et al.
Published: (2025)
RoomEditor++: A Parameter-Sharing Diffusion Architecture for High-Fidelity Furniture Synthesis
by: Wang, Qilong, et al.
Published: (2025)
by: Wang, Qilong, et al.
Published: (2025)
Efficient Masked AutoEncoder for Video Object Counting and A Large-Scale Benchmark
by: Cao, Bing, et al.
Published: (2024)
by: Cao, Bing, et al.
Published: (2024)
Motion-Aware Adaptive Pixel Pruning for Efficient Local Motion Deblurring
by: Shang, Wei, et al.
Published: (2025)
by: Shang, Wei, et al.
Published: (2025)
TAMT: Temporal-Aware Model Tuning for Cross-Domain Few-Shot Action Recognition
by: Wang, Yilong, et al.
Published: (2024)
by: Wang, Yilong, et al.
Published: (2024)
Constrained Prompt Enhancement for Improving Zero-Shot Generalization of Vision-Language Models
by: Yin, Xiaojie, et al.
Published: (2025)
by: Yin, Xiaojie, et al.
Published: (2025)
RGBX-R1: Visual Modality Chain-of-Thought Guided Reinforcement Learning for Multimodal Grounding
by: Wu, Jiahe, et al.
Published: (2026)
by: Wu, Jiahe, et al.
Published: (2026)
Generative Inbetweening through Frame-wise Conditions-Driven Video Generation
by: Zhu, Tianyi, et al.
Published: (2024)
by: Zhu, Tianyi, et al.
Published: (2024)
BackMix: Regularizing Open Set Recognition by Removing Underlying Fore-Background Priors
by: Wang, Yu, et al.
Published: (2025)
by: Wang, Yu, et al.
Published: (2025)
SelfDRSC++: Self-Supervised Learning for Dual Reversed Rolling Shutter Correction
by: Shang, Wei, et al.
Published: (2024)
by: Shang, Wei, et al.
Published: (2024)
BasicAVSR: Arbitrary-Scale Video Super-Resolution via Image Priors and Enhanced Motion Compensation
by: Shang, Wei, et al.
Published: (2025)
by: Shang, Wei, et al.
Published: (2025)
High-Frequency Prior-Driven Adaptive Masking for Accelerating Image Super-Resolution
by: Shang, Wei, et al.
Published: (2025)
by: Shang, Wei, et al.
Published: (2025)
$S^3$: Synonymous Semantic Space for Improving Zero-Shot Generalization of Vision-Language Models
by: Yin, Xiaojie, et al.
Published: (2024)
by: Yin, Xiaojie, et al.
Published: (2024)
DIFF-MF: A Difference-Driven Channel-Spatial State Space Model for Multi-Modal Image Fusion
by: Sun, Yiming, et al.
Published: (2026)
by: Sun, Yiming, et al.
Published: (2026)
Conditional Controllable Image Fusion
by: Cao, Bing, et al.
Published: (2024)
by: Cao, Bing, et al.
Published: (2024)
Reversible Efficient Diffusion for Image Fusion
by: Xu, Xingxin, et al.
Published: (2026)
by: Xu, Xingxin, et al.
Published: (2026)
Asymmetric Reinforcing against Multi-modal Representation Bias
by: Gao, Xiyuan, et al.
Published: (2025)
by: Gao, Xiyuan, et al.
Published: (2025)
A$^2$M$^2$-Net: Adaptively Aligned Multi-Scale Moment for Few-Shot Action Recognition
by: Gao, Zilin, et al.
Published: (2025)
by: Gao, Zilin, et al.
Published: (2025)
VTD-CLIP: Video-to-Text Discretization via Prompting CLIP
by: Zhu, Wencheng, et al.
Published: (2025)
by: Zhu, Wencheng, et al.
Published: (2025)
Exploring Diverse Representations for Open Set Recognition
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
CKD: Contrastive Knowledge Distillation from A Sample-wise Perspective
by: Zhu, Wencheng, et al.
Published: (2024)
by: Zhu, Wencheng, et al.
Published: (2024)
Fine-Grained Domain Generalization with Feature Structuralization
by: Yu, Wenlong, et al.
Published: (2024)
by: Yu, Wenlong, et al.
Published: (2024)
Visible and Clear: Finding Tiny Objects in Difference Map
by: Cao, Bing, et al.
Published: (2024)
by: Cao, Bing, et al.
Published: (2024)
Task-Customized Mixture of Adapters for General Image Fusion
by: Zhu, Pengfei, et al.
Published: (2024)
by: Zhu, Pengfei, et al.
Published: (2024)
CtrlFuse: Mask-Prompt Guided Controllable Infrared and Visible Image Fusion
by: Sun, Yiming, et al.
Published: (2026)
by: Sun, Yiming, et al.
Published: (2026)
Dynamic Brightness Adaptation for Robust Multi-modal Image Fusion
by: Sun, Yiming, et al.
Published: (2024)
by: Sun, Yiming, et al.
Published: (2024)
CoE: Chain-of-Explanation via Automatic Visual Concept Circuit Description and Polysemanticity Quantification
by: Yu, Wenlong, et al.
Published: (2025)
by: Yu, Wenlong, et al.
Published: (2025)
Dream-IF: Dynamic Relative EnhAnceMent for Image Fusion
by: Xu, Xingxin, et al.
Published: (2025)
by: Xu, Xingxin, et al.
Published: (2025)
Bi-directional Self-Registration for Misaligned Infrared-Visible Image Fusion
by: Li, Timing, et al.
Published: (2025)
by: Li, Timing, et al.
Published: (2025)
AutoIAD: Manager-Driven Multi-Agent Collaboration for Automated Industrial Anomaly Detection
by: Ji, Dongwei, et al.
Published: (2025)
by: Ji, Dongwei, et al.
Published: (2025)
Multiple-Exit Tuning: Towards Inference-Efficient Adaptation for Vision Transformer
by: Liu, Zheng, et al.
Published: (2024)
by: Liu, Zheng, et al.
Published: (2024)
Dynamic Sub-graph Distillation for Robust Semi-supervised Continual Learning
by: Fan, Yan, et al.
Published: (2023)
by: Fan, Yan, et al.
Published: (2023)
Multi-view Deep Subspace Clustering Networks
by: Zhu, Pengfei, et al.
Published: (2019)
by: Zhu, Pengfei, et al.
Published: (2019)
DALIP: Distribution Alignment-based Language-Image Pre-Training for Domain-Specific Data
by: Wu, Junjie, et al.
Published: (2025)
by: Wu, Junjie, et al.
Published: (2025)
Efficient-VLN: A Training-Efficient Vision-Language Navigation Model
by: Zheng, Duo, et al.
Published: (2025)
by: Zheng, Duo, et al.
Published: (2025)
Thin-Plate Spline-based Interpolation for Animation Line Inbetweening
by: Zhu, Tianyi, et al.
Published: (2024)
by: Zhu, Tianyi, et al.
Published: (2024)
PIDNet: Progressive Implicit Decouple Network for Multimodal Action Quality Assessment
by: Li, Qiqi, et al.
Published: (2026)
by: Li, Qiqi, et al.
Published: (2026)
SparseVILA: Decoupling Visual Sparsity for Efficient VLM Inference
by: Khaki, Samir, et al.
Published: (2025)
by: Khaki, Samir, et al.
Published: (2025)
Sparse-Tuning: Adapting Vision Transformers with Efficient Fine-tuning and Inference
by: Liu, Ting, et al.
Published: (2024)
by: Liu, Ting, et al.
Published: (2024)
Similar Items
-
AMU-Tuning: Effective Logit Bias for CLIP-based Few-shot Learning
by: Tang, Yuwei, et al.
Published: (2024) -
Unleashing Degradation-Carrying Features in Symmetric U-Net: Simpler and Stronger Baselines for All-in-One Image Restoration
by: Jiao, Wenlong, et al.
Published: (2025) -
RoomEditor++: A Parameter-Sharing Diffusion Architecture for High-Fidelity Furniture Synthesis
by: Wang, Qilong, et al.
Published: (2025) -
Efficient Masked AutoEncoder for Video Object Counting and A Large-Scale Benchmark
by: Cao, Bing, et al.
Published: (2024) -
Motion-Aware Adaptive Pixel Pruning for Efficient Local Motion Deblurring
by: Shang, Wei, et al.
Published: (2025)