ProAPO: Progressively Automatic Prompt Optimization for Visual Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Qu, Xiangyan, Gou, Gaopeng, Zhuang, Jiamin, Yu, Jing, Song, Kun, Wang, Qihao, Li, Yili, Xiong, Gang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MADS: Multi-Attribute Document Supervision for Zero-Shot Image Classification
by: Qu, Xiangyan, et al.
Published: (2025)
by: Qu, Xiangyan, et al.
Published: (2025)
T2VParser: Adaptive Decomposition Tokens for Partial Alignment in Text to Video Retrieval
by: Li, Yili, et al.
Published: (2025)
by: Li, Yili, et al.
Published: (2025)
Visual-Semantic Decomposition and Partial Alignment for Document-based Zero-Shot Learning
by: Qu, Xiangyan, et al.
Published: (2024)
by: Qu, Xiangyan, et al.
Published: (2024)
Missing Target-Relevant Information Prediction with World Model for Accurate Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2025)
by: Tang, Yuanmin, et al.
Published: (2025)
Denoise-I2W: Mapping Images to Denoising Words for Accurate Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2024)
by: Tang, Yuanmin, et al.
Published: (2024)
UniAPO: Unified Multimodal Automated Prompt Optimization
by: Zhu, Qipeng, et al.
Published: (2025)
by: Zhu, Qipeng, et al.
Published: (2025)
IIU: Independent Inference Units for Knowledge-based Visual Question Answering
by: Li, Yili, et al.
Published: (2024)
by: Li, Yili, et al.
Published: (2024)
From Scale to Speed: Adaptive Test-Time Scaling for Image Editing
by: Qu, Xiangyan, et al.
Published: (2026)
by: Qu, Xiangyan, et al.
Published: (2026)
ProEdit: Simple Progression is All You Need for High-Quality 3D Scene Editing
by: Chen, Jun-Kun, et al.
Published: (2024)
by: Chen, Jun-Kun, et al.
Published: (2024)
Reason-before-Retrieve: One-Stage Reflective Chain-of-Thoughts for Training-Free Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2024)
by: Tang, Yuanmin, et al.
Published: (2024)
T2VIndexer: A Generative Video Indexer for Efficient Text-Video Retrieval
by: Li, Yili, et al.
Published: (2024)
by: Li, Yili, et al.
Published: (2024)
Neuro-3D: Towards 3D Visual Decoding from EEG Signals
by: Guo, Zhanqiang, et al.
Published: (2024)
by: Guo, Zhanqiang, et al.
Published: (2024)
APO: Enhancing Reasoning Ability of MLLMs via Asymmetric Policy Optimization
by: Hong, Minjie, et al.
Published: (2025)
by: Hong, Minjie, et al.
Published: (2025)
ProTeCt: Prompt Tuning for Taxonomic Open Set Classification
by: Wu, Tz-Ying, et al.
Published: (2023)
by: Wu, Tz-Ying, et al.
Published: (2023)
ProEdit: Inversion-based Editing From Prompts Done Right
by: Ouyang, Zhi, et al.
Published: (2025)
by: Ouyang, Zhi, et al.
Published: (2025)
Token Coordinated Prompt Attention is Needed for Visual Prompting
by: Liu, Zichen, et al.
Published: (2025)
by: Liu, Zichen, et al.
Published: (2025)
Multi-Prompt Progressive Alignment for Multi-Source Unsupervised Domain Adaptation
by: Chen, Haoran, et al.
Published: (2025)
by: Chen, Haoran, et al.
Published: (2025)
Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers
by: Zhu, Jingyuan, et al.
Published: (2026)
by: Zhu, Jingyuan, et al.
Published: (2026)
Think Twice to See More: Iterative Visual Reasoning in Medical VLMs
by: Chen, Kaitao, et al.
Published: (2025)
by: Chen, Kaitao, et al.
Published: (2025)
ProDehaze: Prompting Diffusion Models Toward Faithful Image Dehazing
by: Zhou, Tianwen, et al.
Published: (2025)
by: Zhou, Tianwen, et al.
Published: (2025)
VisualPrompter: Semantic-Aware Prompt Optimization with Visual Feedback for Text-to-Image Synthesis
by: Wu, Shiyu, et al.
Published: (2025)
by: Wu, Shiyu, et al.
Published: (2025)
Retrieval-Enhanced Visual Prompt Learning for Few-shot Classification
by: Rong, Jintao, et al.
Published: (2023)
by: Rong, Jintao, et al.
Published: (2023)
Visual Prompt-Agnostic Evolution
by: Wang, Junze, et al.
Published: (2026)
by: Wang, Junze, et al.
Published: (2026)
ProVG: Progressive Visual Grounding via Language Decoupling for Remote Sensing Imagery
by: Li, Ke, et al.
Published: (2026)
by: Li, Ke, et al.
Published: (2026)
MindAligner: Explicit Brain Functional Alignment for Cross-Subject Visual Decoding from Limited fMRI Data
by: Dai, Yuqin, et al.
Published: (2025)
by: Dai, Yuqin, et al.
Published: (2025)
FedAPT: Federated Adversarial Prompt Tuning for Vision-Language Models
by: Zhai, Kun, et al.
Published: (2025)
by: Zhai, Kun, et al.
Published: (2025)
ProSAM: Enhancing the Robustness of SAM-based Visual Reference Segmentation with Probabilistic Prompts
by: Wang, Xiaoqi, et al.
Published: (2025)
by: Wang, Xiaoqi, et al.
Published: (2025)
ProPhy: Progressive Physical Alignment for Dynamic World Simulation
by: Wang, Zijun, et al.
Published: (2025)
by: Wang, Zijun, et al.
Published: (2025)
Progressive Visual Prompt Learning with Contrastive Feature Re-formation
by: Xu, Chen, et al.
Published: (2023)
by: Xu, Chen, et al.
Published: (2023)
MMSD3.0: A Multi-Image Benchmark for Real-World Multimodal Sarcasm Detection
by: Zhao, Haochen, et al.
Published: (2025)
by: Zhao, Haochen, et al.
Published: (2025)
AdaViPro: Region-based Adaptive Visual Prompt for Large-Scale Models Adapting
by: Yang, Mengyu, et al.
Published: (2024)
by: Yang, Mengyu, et al.
Published: (2024)
Efficient Multi-Instance Generation with Janus-Pro-Dirven Prompt Parsing
by: Qi, Fan, et al.
Published: (2025)
by: Qi, Fan, et al.
Published: (2025)
SynBrain: Enhancing Visual-to-fMRI Synthesis via Probabilistic Representation Learning
by: Mai, Weijian, et al.
Published: (2025)
by: Mai, Weijian, et al.
Published: (2025)
ProS: Prompting-to-simulate Generalized knowledge for Universal Cross-Domain Retrieval
by: Fang, Kaipeng, et al.
Published: (2023)
by: Fang, Kaipeng, et al.
Published: (2023)
Enhancing Image Generation Fidelity via Progressive Prompts
by: Xiong, Zhen, et al.
Published: (2025)
by: Xiong, Zhen, et al.
Published: (2025)
ProJo4D: Progressive Joint Optimization for Sparse-View Inverse Physics Estimation
by: Rho, Daniel, et al.
Published: (2025)
by: Rho, Daniel, et al.
Published: (2025)
ProMamba: Prompt-Mamba for polyp segmentation
by: Xie, Jianhao, et al.
Published: (2024)
by: Xie, Jianhao, et al.
Published: (2024)
Solving Zero-Shot 3D Visual Grounding as Constraint Satisfaction Problems
by: Yuan, Qihao, et al.
Published: (2024)
by: Yuan, Qihao, et al.
Published: (2024)
Multi-Scale Visual Prompting for Lightweight Small-Image Classification
by: Khazem, Salim
Published: (2025)
by: Khazem, Salim
Published: (2025)
Beyond Random: Automatic Inner-loop Optimization in Dataset Distillation
by: Li, Muquan, et al.
Published: (2025)
by: Li, Muquan, et al.
Published: (2025)
Similar Items
-
MADS: Multi-Attribute Document Supervision for Zero-Shot Image Classification
by: Qu, Xiangyan, et al.
Published: (2025) -
T2VParser: Adaptive Decomposition Tokens for Partial Alignment in Text to Video Retrieval
by: Li, Yili, et al.
Published: (2025) -
Visual-Semantic Decomposition and Partial Alignment for Document-based Zero-Shot Learning
by: Qu, Xiangyan, et al.
Published: (2024) -
Missing Target-Relevant Information Prediction with World Model for Accurate Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2025) -
Denoise-I2W: Mapping Images to Denoising Words for Accurate Zero-Shot Composed Image Retrieval
by: Tang, Yuanmin, et al.
Published: (2024)