From Generation to Generalization: Emergent Few-Shot Learning in Video Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Acuaviva, Pablo, Davtyan, Aram, Hassan, Mariam, Stapf, Sebastian, Rahimi, Ahmad, Alahi, Alexandre, Favaro, Paolo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Rethinking Visual Intelligence: Insights from Video Pretraining
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
Communication-Inspired Tokenization for Structured Image Representations
von: Davtyan, Aram, et al.
Veröffentlicht: (2026)
von: Davtyan, Aram, et al.
Veröffentlicht: (2026)
Learn the Force We Can: Enabling Sparse Motion Control in Multi-Object Video Generation
von: Davtyan, Aram, et al.
Veröffentlicht: (2023)
von: Davtyan, Aram, et al.
Veröffentlicht: (2023)
Composition of Memory Experts for Diffusion World Models
von: Stapf, Sebastian, et al.
Veröffentlicht: (2026)
von: Stapf, Sebastian, et al.
Veröffentlicht: (2026)
CAGE: Unsupervised Visual Composition and Animation for Controllable Video Generation
von: Davtyan, Aram, et al.
Veröffentlicht: (2024)
von: Davtyan, Aram, et al.
Veröffentlicht: (2024)
Anchored Video Generation: Decoupling Scene Construction and Temporal Synthesis in Text-to-Video Diffusion Models
von: Hassan, Mariam, et al.
Veröffentlicht: (2025)
von: Hassan, Mariam, et al.
Veröffentlicht: (2025)
Faster Inference of Flow-Based Generative Models via Improved Data-Noise Coupling
von: Davtyan, Aram, et al.
Veröffentlicht: (2026)
von: Davtyan, Aram, et al.
Veröffentlicht: (2026)
Proprio: Latent Self-Scoring and Inference-Time Refinement for Physically Plausible Video Generation
von: Hassan, Mariam, et al.
Veröffentlicht: (2026)
von: Hassan, Mariam, et al.
Veröffentlicht: (2026)
GEM: A Generalizable Ego-Vision Multimodal World Model for Fine-Grained Ego-Motion, Object Dynamics, and Scene Composition Control
von: Hassan, Mariam, et al.
Veröffentlicht: (2024)
von: Hassan, Mariam, et al.
Veröffentlicht: (2024)
A Multi-Loss Strategy for Vehicle Trajectory Prediction: Combining Off-Road, Diversity, and Directional Consistency Losses
von: Rahimi, Ahmad, et al.
Veröffentlicht: (2024)
von: Rahimi, Ahmad, et al.
Veröffentlicht: (2024)
KOALA: A Kalman Optimization Algorithm with Loss Adaptivity
von: Davtyan, Aram, et al.
Veröffentlicht: (2021)
von: Davtyan, Aram, et al.
Veröffentlicht: (2021)
MAD: Motion Appearance Decoupling for efficient Driving World Models
von: Rahimi, Ahmad, et al.
Veröffentlicht: (2026)
von: Rahimi, Ahmad, et al.
Veröffentlicht: (2026)
LayerSync: Self-aligning Intermediate Layers
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2025)
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2025)
Co-Supervised Learning: Improving Weak-to-Strong Generalization with Hierarchical Mixture of Experts
von: Liu, Yuejiang, et al.
Veröffentlicht: (2024)
von: Liu, Yuejiang, et al.
Veröffentlicht: (2024)
SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2026)
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2026)
AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
von: Hu, Teng, et al.
Veröffentlicht: (2023)
von: Hu, Teng, et al.
Veröffentlicht: (2023)
Stable Video Infinity: Infinite-Length Video Generation with Error Recycling
von: Li, Wuyang, et al.
Veröffentlicht: (2025)
von: Li, Wuyang, et al.
Veröffentlicht: (2025)
Invert2Restore: Zero-Shot Degradation-Blind Image Restoration
von: Chihaoui, Hamadi, et al.
Veröffentlicht: (2025)
von: Chihaoui, Hamadi, et al.
Veröffentlicht: (2025)
Diffusion Image Prior
von: Chihaoui, Hamadi, et al.
Veröffentlicht: (2025)
von: Chihaoui, Hamadi, et al.
Veröffentlicht: (2025)
Dual-Interrelated Diffusion Model for Few-Shot Anomaly Image Generation
von: Jin, Ying, et al.
Veröffentlicht: (2024)
von: Jin, Ying, et al.
Veröffentlicht: (2024)
KOALA++: Efficient Kalman-Based Optimization with Gradient-Covariance Products
von: Xia, Zixuan, et al.
Veröffentlicht: (2025)
von: Xia, Zixuan, et al.
Veröffentlicht: (2025)
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models
von: Kara, Ozgur, et al.
Veröffentlicht: (2025)
von: Kara, Ozgur, et al.
Veröffentlicht: (2025)
Sim-to-Real Causal Transfer: A Metric Learning Approach to Causally-Aware Interaction Representations
von: Rahimi, Ahmad, et al.
Veröffentlicht: (2023)
von: Rahimi, Ahmad, et al.
Veröffentlicht: (2023)
Few-Shot Synthetic Data Generation with Diffusion Models for Downstream Vision Tasks
von: Dushenev, Daniil, et al.
Veröffentlicht: (2026)
von: Dushenev, Daniil, et al.
Veröffentlicht: (2026)
Few-Shot-Based Modular Image-to-Video Adapter for Diffusion Models
von: Li, Zhenhao, et al.
Veröffentlicht: (2025)
von: Li, Zhenhao, et al.
Veröffentlicht: (2025)
A Feature Generator for Few-Shot Learning
von: Kanagalingam, Heethanjan, et al.
Veröffentlicht: (2024)
von: Kanagalingam, Heethanjan, et al.
Veröffentlicht: (2024)
Few-Shot Image Generation by Conditional Relaxing Diffusion Inversion
von: Cao, Yu, et al.
Veröffentlicht: (2024)
von: Cao, Yu, et al.
Veröffentlicht: (2024)
NAT: Learning to Attack Neurons for Enhanced Adversarial Transferability
von: Nakka, Krishna Kanth, et al.
Veröffentlicht: (2025)
von: Nakka, Krishna Kanth, et al.
Veröffentlicht: (2025)
Zero-Shot Video Deraining with Video Diffusion Models
von: Varanka, Tuomas, et al.
Veröffentlicht: (2025)
von: Varanka, Tuomas, et al.
Veröffentlicht: (2025)
FlowCut: Unsupervised Video Instance Segmentation via Temporal Mask Matching
von: Sari, Alp Eren, et al.
Veröffentlicht: (2025)
von: Sari, Alp Eren, et al.
Veröffentlicht: (2025)
Few-Part-Shot Font Generation
von: Akiba, Masaki, et al.
Veröffentlicht: (2025)
von: Akiba, Masaki, et al.
Veröffentlicht: (2025)
FG$^2$: Fine-Grained Cross-View Localization by Fine-Grained Feature Matching
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
Edit2Interp: Adapting Image Foundation Models from Spatial Editing to Video Frame Interpolation with Few-Shot Learning
von: Rahimi, Nasrin, et al.
Veröffentlicht: (2026)
von: Rahimi, Nasrin, et al.
Veröffentlicht: (2026)
UNEM: UNrolled Generalized EM for Transductive Few-Shot Learning
von: Zhou, Long, et al.
Veröffentlicht: (2024)
von: Zhou, Long, et al.
Veröffentlicht: (2024)
Decomposed Prototype Learning for Few-Shot Scene Graph Generation
von: Li, Xingchen, et al.
Veröffentlicht: (2023)
von: Li, Xingchen, et al.
Veröffentlicht: (2023)
Blind Image Restoration via Fast Diffusion Inversion
von: Chihaoui, Hamadi, et al.
Veröffentlicht: (2024)
von: Chihaoui, Hamadi, et al.
Veröffentlicht: (2024)
Towards Generalizable Trajectory Prediction Using Dual-Level Representation Learning And Adaptive Prompting
von: Messaoud, Kaouther, et al.
Veröffentlicht: (2025)
von: Messaoud, Kaouther, et al.
Veröffentlicht: (2025)
Structure-Level Disentangled Diffusion for Few-Shot Chinese Font Generation
von: Li, Jie, et al.
Veröffentlicht: (2026)
von: Li, Jie, et al.
Veröffentlicht: (2026)
SWIFT: Sliding Window Reconstruction for Few-Shot Training-Free Generated Video Attribution
von: Wang, Chao, et al.
Veröffentlicht: (2026)
von: Wang, Chao, et al.
Veröffentlicht: (2026)
Unlocking Compositional Generalization in Continual Few-Shot Learning
von: Nguyen-Lam, Phu-Quy, et al.
Veröffentlicht: (2026)
von: Nguyen-Lam, Phu-Quy, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Rethinking Visual Intelligence: Insights from Video Pretraining
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025) -
Communication-Inspired Tokenization for Structured Image Representations
von: Davtyan, Aram, et al.
Veröffentlicht: (2026) -
Learn the Force We Can: Enabling Sparse Motion Control in Multi-Object Video Generation
von: Davtyan, Aram, et al.
Veröffentlicht: (2023) -
Composition of Memory Experts for Diffusion World Models
von: Stapf, Sebastian, et al.
Veröffentlicht: (2026) -
CAGE: Unsupervised Visual Composition and Animation for Controllable Video Generation
von: Davtyan, Aram, et al.
Veröffentlicht: (2024)