I2V-Adapter: A General Image-to-Video Adapter for Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Guo, Xun, Zheng, Mingwu, Hou, Liang, Gao, Yuan, Deng, Yufan, Wan, Pengfei, Zhang, Di, Liu, Yufan, Hu, Weiming, Zha, Zhengjun, Huang, Haibin, Ma, Chongyang |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diffusion Restoration Adapter for Real-World Image Restoration
by: Liang, Hanbang, et al.
Published: (2025)
by: Liang, Hanbang, et al.
Published: (2025)
VRMM: A Volumetric Relightable Morphable Head Model
by: Yang, Haotian, et al.
Published: (2024)
by: Yang, Haotian, et al.
Published: (2024)
CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance
by: Deng, Yufan, et al.
Published: (2025)
by: Deng, Yufan, et al.
Published: (2025)
Few-Shot-Based Modular Image-to-Video Adapter for Diffusion Models
by: Li, Zhenhao, et al.
Published: (2025)
by: Li, Zhenhao, et al.
Published: (2025)
Direct-a-Video: Customized Video Generation with User-Directed Camera Movement and Object Motion
by: Yang, Shiyuan, et al.
Published: (2024)
by: Yang, Shiyuan, et al.
Published: (2024)
MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement
by: Deng, Yufan, et al.
Published: (2025)
by: Deng, Yufan, et al.
Published: (2025)
Task-Customized Mixture of Adapters for General Image Fusion
by: Zhu, Pengfei, et al.
Published: (2024)
by: Zhu, Pengfei, et al.
Published: (2024)
Boosting Resolution Generalization of Diffusion Transformers with Randomized Positional Encodings
by: Hou, Liang, et al.
Published: (2025)
by: Hou, Liang, et al.
Published: (2025)
ResAdapter: Domain Consistent Resolution Adapter for Diffusion Models
by: Cheng, Jiaxiang, et al.
Published: (2024)
by: Cheng, Jiaxiang, et al.
Published: (2024)
Beyond Dataset Distillation: Lossless Dataset Concentration via Diffusion-Assisted Distribution Alignment
by: Liu, Tongfei, et al.
Published: (2026)
by: Liu, Tongfei, et al.
Published: (2026)
FE-Adapter: Adapting Image-based Emotion Classifiers to Videos
by: Gowda, Shreyank N, et al.
Published: (2024)
by: Gowda, Shreyank N, et al.
Published: (2024)
HumanNet: Scaling Human-centric Video Learning to One Million Hours
by: Deng, Yufan, et al.
Published: (2026)
by: Deng, Yufan, et al.
Published: (2026)
PFSD: A Multi-Modal Pedestrian-Focus Scene Dataset for Rich Tasks in Semi-Structured Environments
by: Liu, Yueting, et al.
Published: (2025)
by: Liu, Yueting, et al.
Published: (2025)
Adapter-dependent Adapter Methylation Assay
by: Zhang, Jia, et al.
Published: (2024)
by: Zhang, Jia, et al.
Published: (2024)
UNIC-Adapter: Unified Image-instruction Adapter with Multi-modal Transformer for Image Generation
by: Duan, Lunhao, et al.
Published: (2024)
by: Duan, Lunhao, et al.
Published: (2024)
SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
by: Shen, Guibao, et al.
Published: (2024)
by: Shen, Guibao, et al.
Published: (2024)
Mod-Adapter: Tuning-Free and Versatile Multi-concept Personalization via Modulation Adapter
by: Zhong, Weizhi, et al.
Published: (2025)
by: Zhong, Weizhi, et al.
Published: (2025)
Dual-level Adaptation for Multi-Object Tracking: Building Test-Time Calibration from Experience and Intuition
by: Guo, Wen, et al.
Published: (2026)
by: Guo, Wen, et al.
Published: (2026)
CLIP-Adapter: Better Vision-Language Models with Feature Adapters
by: Gao, Peng, et al.
Published: (2021)
by: Gao, Peng, et al.
Published: (2021)
Q-Adapter: Visual Query Adapter for Extracting Textually-related Features in Video Captioning
by: Chen, Junan, et al.
Published: (2025)
by: Chen, Junan, et al.
Published: (2025)
VideoTetris: Towards Compositional Text-to-Video Generation
by: Tian, Ye, et al.
Published: (2024)
by: Tian, Ye, et al.
Published: (2024)
Motion-Adapter: A Diffusion Model Adapter for Text-to-Motion Generation of Compound Actions
by: Jiang, Yue, et al.
Published: (2026)
by: Jiang, Yue, et al.
Published: (2026)
RAP: Efficient Text-Video Retrieval with Sparse-and-Correlated Adapter
by: Cao, Meng, et al.
Published: (2024)
by: Cao, Meng, et al.
Published: (2024)
Inv-Adapter: ID Customization Generation via Image Inversion and Lightweight Adapter
by: Xing, Peng, et al.
Published: (2024)
by: Xing, Peng, et al.
Published: (2024)
Hyperspectral Adapter for Object Tracking based on Hyperspectral Video
by: Gao, Long, et al.
Published: (2025)
by: Gao, Long, et al.
Published: (2025)
CAT: Contrastive Adapter Training for Personalized Image Generation
by: Park, Jae Wan, et al.
Published: (2024)
by: Park, Jae Wan, et al.
Published: (2024)
Visual-Instructed Degradation Diffusion for All-in-One Image Restoration
by: Luo, Wenyang, et al.
Published: (2025)
by: Luo, Wenyang, et al.
Published: (2025)
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models
by: Kara, Ozgur, et al.
Published: (2025)
by: Kara, Ozgur, et al.
Published: (2025)
Focal Guidance: Unlocking Controllability from Semantic-Weak Layers in Video Diffusion Models
by: Yin, Yuanyang, et al.
Published: (2026)
by: Yin, Yuanyang, et al.
Published: (2026)
MV-Adapter: Multimodal Video Transfer Learning for Video Text Retrieval
by: Jin, Xiaojie, et al.
Published: (2023)
by: Jin, Xiaojie, et al.
Published: (2023)
Att-Adapter: A Robust and Precise Domain-Specific Multi-Attributes T2I Diffusion Adapter via Conditional Variational Autoencoder
by: Cho, Wonwoong, et al.
Published: (2025)
by: Cho, Wonwoong, et al.
Published: (2025)
The Illusion of Forgetting: Attack Unlearned Diffusion via Initial Latent Variable Optimization
by: Li, Manyi, et al.
Published: (2026)
by: Li, Manyi, et al.
Published: (2026)
LD-LAudio-V1: Video-to-Long-Form-Audio Generation Extension with Dual Lightweight Adapters
by: Zhang, Haomin, et al.
Published: (2025)
by: Zhang, Haomin, et al.
Published: (2025)
Instant Video Models: Universal Adapters for Stabilizing Image-Based Networks
by: Dutson, Matthew, et al.
Published: (2025)
by: Dutson, Matthew, et al.
Published: (2025)
LQ-Adapter: ViT-Adapter with Learnable Queries for Gallbladder Cancer Detection from Ultrasound Image
by: Madan, Chetan, et al.
Published: (2024)
by: Madan, Chetan, et al.
Published: (2024)
Causal-Adapter: Taming Text-to-Image Diffusion for Faithful Counterfactual Generation
by: Tong, Lei, et al.
Published: (2025)
by: Tong, Lei, et al.
Published: (2025)
Generalizing Deepfake Video Detection with Plug-and-Play: Video-Level Blending and Spatiotemporal Adapter Tuning
by: Yan, Zhiyuan, et al.
Published: (2024)
by: Yan, Zhiyuan, et al.
Published: (2024)
MV-Adapter: Multi-view Consistent Image Generation Made Easy
by: Huang, Zehuan, et al.
Published: (2024)
by: Huang, Zehuan, et al.
Published: (2024)
Adapter-Enhanced Semantic Prompting for Continual Learning
by: Yin, Baocai, et al.
Published: (2024)
by: Yin, Baocai, et al.
Published: (2024)
Federated Client-tailored Adapter for Medical Image Segmentation
by: Hu, Guyue, et al.
Published: (2025)
by: Hu, Guyue, et al.
Published: (2025)
Similar Items
-
Diffusion Restoration Adapter for Real-World Image Restoration
by: Liang, Hanbang, et al.
Published: (2025) -
VRMM: A Volumetric Relightable Morphable Head Model
by: Yang, Haotian, et al.
Published: (2024) -
CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance
by: Deng, Yufan, et al.
Published: (2025) -
Few-Shot-Based Modular Image-to-Video Adapter for Diffusion Models
by: Li, Zhenhao, et al.
Published: (2025) -
Direct-a-Video: Customized Video Generation with User-Directed Camera Movement and Object Motion
by: Yang, Shiyuan, et al.
Published: (2024)