Foundation Cures Personalization: Improving Personalized Models' Prompt Consistency via Hidden Foundation Knowledge
Fuente:
arXiv
Saved in:
| Main Authors: | Cai, Yiyang, Jiang, Zhengkai, Liu, Yulong, Jiang, Chunyang, Xue, Wei, Guo, Yike, Luo, Wenhan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PointSeg: A Training-Free Paradigm for 3D Scene Segmentation via Foundation Models
by: He, Qingdong, et al.
Published: (2024)
by: He, Qingdong, et al.
Published: (2024)
Towards Training-free Open-world Segmentation via Image Prompt Foundation Models
by: Tang, Lv, et al.
Published: (2023)
by: Tang, Lv, et al.
Published: (2023)
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model
by: Li, Yan, et al.
Published: (2024)
by: Li, Yan, et al.
Published: (2024)
Backdoor Attacks on Prompt-Driven Video Segmentation Foundation Models
by: Zhang, Zongmin, et al.
Published: (2025)
by: Zhang, Zongmin, et al.
Published: (2025)
HiPrompt: Tuning-free Higher-Resolution Generation with Hierarchical MLLM Prompts
by: Liu, Xinyu, et al.
Published: (2024)
by: Liu, Xinyu, et al.
Published: (2024)
Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
by: Kong, Zhe, et al.
Published: (2025)
by: Kong, Zhe, et al.
Published: (2025)
CogniEdit: Dense Gradient Flow Optimization for Fine-Grained Image Editing
by: Li, Yan, et al.
Published: (2025)
by: Li, Yan, et al.
Published: (2025)
VFX Creator: Animated Visual Effect Generation with Controllable Diffusion Transformer
by: Liu, Xinyu, et al.
Published: (2025)
by: Liu, Xinyu, et al.
Published: (2025)
Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models
by: Wang, Zhiqiang, et al.
Published: (2026)
by: Wang, Zhiqiang, et al.
Published: (2026)
Curriculum Prompting Foundation Models for Medical Image Segmentation
by: Zheng, Xiuqi, et al.
Published: (2024)
by: Zheng, Xiuqi, et al.
Published: (2024)
OMG: Occlusion-friendly Personalized Multi-concept Generation in Diffusion Models
by: Kong, Zhe, et al.
Published: (2024)
by: Kong, Zhe, et al.
Published: (2024)
Fine-Tuning Impairs the Balancedness of Foundation Models in Long-tailed Personalized Federated Learning
by: Hou, Shihao, et al.
Published: (2026)
by: Hou, Shihao, et al.
Published: (2026)
Surgical Depth Anything: Depth Estimation for Surgical Scenes using Foundation Models
by: Lou, Ange, et al.
Published: (2024)
by: Lou, Ange, et al.
Published: (2024)
Learning How To Ask: Cycle-Consistency Refines Prompts in Multimodal Foundation Models
by: Diesendruck, Maurice, et al.
Published: (2024)
by: Diesendruck, Maurice, et al.
Published: (2024)
A Survey on Foundation Models for Personalized Federated Intelligence
by: Qiao, Yu, et al.
Published: (2025)
by: Qiao, Yu, et al.
Published: (2025)
Benchmarking Pathology Foundation Models for Spatial Domain Understanding
by: Zhao, Bokai, et al.
Published: (2026)
by: Zhao, Bokai, et al.
Published: (2026)
Personalized Federated Fine-Tuning of Vision Foundation Models for Healthcare
by: Tupper, Adam, et al.
Published: (2025)
by: Tupper, Adam, et al.
Published: (2025)
OSV: One Step is Enough for High-Quality Image to Video Generation
by: Mao, Xiaofeng, et al.
Published: (2024)
by: Mao, Xiaofeng, et al.
Published: (2024)
Prompting Continual Person Search
by: Zhang, Pengcheng, et al.
Published: (2024)
by: Zhang, Pengcheng, et al.
Published: (2024)
A Token-level Text Image Foundation Model for Document Understanding
by: Guan, Tongkun, et al.
Published: (2025)
by: Guan, Tongkun, et al.
Published: (2025)
Tell2Adapt: A Unified Framework for Source Free Unsupervised Domain Adaptation via Vision Foundation Model
by: Shi, Yulong, et al.
Published: (2026)
by: Shi, Yulong, et al.
Published: (2026)
Supervised Fine-tuning in turn Improves Visual Foundation Models
by: Jiang, Xiaohu, et al.
Published: (2024)
by: Jiang, Xiaohu, et al.
Published: (2024)
Online3R: Online Learning for Consistent Sequential Reconstruction Based on Geometry Foundation Model
by: Zhou, Shunkai, et al.
Published: (2026)
by: Zhou, Shunkai, et al.
Published: (2026)
SapiensID: Foundation for Human Recognition
by: Kim, Minchul, et al.
Published: (2025)
by: Kim, Minchul, et al.
Published: (2025)
Personal Visual Context Learning in Large Multimodal Models
by: Xue, Zihui, et al.
Published: (2026)
by: Xue, Zihui, et al.
Published: (2026)
An Improved Method for Personalizing Diffusion Models
by: Zeng, Yan, et al.
Published: (2024)
by: Zeng, Yan, et al.
Published: (2024)
RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency
by: Li, Siqi, et al.
Published: (2025)
by: Li, Siqi, et al.
Published: (2025)
Multi-View Large Reconstruction Model via Geometry-Aware Positional Encoding and Attention
by: Li, Mengfei, et al.
Published: (2024)
by: Li, Mengfei, et al.
Published: (2024)
Co$^{3}$Gesture: Towards Coherent Concurrent Co-speech 3D Gesture Generation with Interactive Diffusion
by: Qi, Xingqun, et al.
Published: (2025)
by: Qi, Xingqun, et al.
Published: (2025)
DEFOM-Stereo: Depth Foundation Model Based Stereo Matching
by: Jiang, Hualie, et al.
Published: (2025)
by: Jiang, Hualie, et al.
Published: (2025)
TALO: Pushing 3D Vision Foundation Models Towards Globally Consistent Online Reconstruction
by: Zhang, Fengyi, et al.
Published: (2025)
by: Zhang, Fengyi, et al.
Published: (2025)
Genome-Anchored Foundation Model Embeddings Improve Molecular Prediction from Histology Images
by: Jin, Cheng, et al.
Published: (2025)
by: Jin, Cheng, et al.
Published: (2025)
Foundation Model-guided Iteratively Prompting and Pseudo-Labeling for Partially Labeled Medical Image Segmentation
by: Zhao, Qiaochu, et al.
Published: (2026)
by: Zhao, Qiaochu, et al.
Published: (2026)
GuiDINO: Rethinking Vision Foundation Model in Medical Image Segmentation
by: Liang, Zhuonan, et al.
Published: (2026)
by: Liang, Zhuonan, et al.
Published: (2026)
Consistent View Alignment Improves Foundation Models for 3D Medical Image Segmentation
by: Vaish, Puru, et al.
Published: (2025)
by: Vaish, Puru, et al.
Published: (2025)
Scene-Adaptive Person Search via Bilateral Modulations
by: Jiang, Yimin, et al.
Published: (2024)
by: Jiang, Yimin, et al.
Published: (2024)
Personalized Federated Learning via Dual-Prompt Optimization and Cross Fusion
by: Zhang, Yuguang, et al.
Published: (2025)
by: Zhang, Yuguang, et al.
Published: (2025)
HPT++: Hierarchically Prompting Vision-Language Models with Multi-Granularity Knowledge Generation and Improved Structure Modeling
by: Wang, Yubin, et al.
Published: (2024)
by: Wang, Yubin, et al.
Published: (2024)
FedBPrompt: Federated Domain Generalization Person Re-Identification via Body Distribution Aware Visual Prompts
by: Xu, Xin, et al.
Published: (2026)
by: Xu, Xin, et al.
Published: (2026)
Simba: Towards High-Fidelity and Geometrically-Consistent Point Cloud Completion via Transformation Diffusion
by: Zhang, Lirui, et al.
Published: (2025)
by: Zhang, Lirui, et al.
Published: (2025)
Similar Items
-
PointSeg: A Training-Free Paradigm for 3D Scene Segmentation via Foundation Models
by: He, Qingdong, et al.
Published: (2024) -
Towards Training-free Open-world Segmentation via Image Prompt Foundation Models
by: Tang, Lv, et al.
Published: (2023) -
SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model
by: Li, Yan, et al.
Published: (2024) -
Backdoor Attacks on Prompt-Driven Video Segmentation Foundation Models
by: Zhang, Zongmin, et al.
Published: (2025) -
HiPrompt: Tuning-free Higher-Resolution Generation with Hierarchical MLLM Prompts
by: Liu, Xinyu, et al.
Published: (2024)