Tuning-free Visual Effect Transfer across Videos
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jones, Maxwell, Abdal, Rameen, Patashnik, Or, Salakhutdinov, Ruslan, Tulyakov, Sergey, Zhu, Jun-Yan, Wang, Kuan-Chieh Jackson |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Visual Personalization Turing Test
von: Abdal, Rameen, et al.
Veröffentlicht: (2026)
von: Abdal, Rameen, et al.
Veröffentlicht: (2026)
Dynamic Concepts Personalization from Single Videos
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
ArtifactLens: Hundreds of Labels Are Enough for Artifact Detection with VLMs
von: Burgess, James, et al.
Veröffentlicht: (2026)
von: Burgess, James, et al.
Veröffentlicht: (2026)
Zero-Shot Dynamic Concept Personalization with Grid-Based LoRA
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
Omni-ID: Holistic Identity Representation Designed for Generative Tasks
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
Helix4D: Complex 4D Mesh Generation
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2026)
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2026)
Improving the Diffusability of Autoencoders
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2025)
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2025)
Interpreting the Weight Space of Customized Diffusion Models
von: Dravid, Amil, et al.
Veröffentlicht: (2024)
von: Dravid, Amil, et al.
Veröffentlicht: (2024)
NearID: Identity Representation Learning via Near-identity Distractors
von: Cvejic, Aleksandar, et al.
Veröffentlicht: (2026)
von: Cvejic, Aleksandar, et al.
Veröffentlicht: (2026)
Object-level Visual Prompts for Compositional Image Generation
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
Nested Attention: Semantic-aware Attention Values for Concept Personalization
von: Patashnik, Or, et al.
Veröffentlicht: (2025)
von: Patashnik, Or, et al.
Veröffentlicht: (2025)
MoA: Mixture-of-Attention for Subject-Context Disentanglement in Personalized Image Generation
von: Wang, Kuan-Chieh, et al.
Veröffentlicht: (2024)
von: Wang, Kuan-Chieh, et al.
Veröffentlicht: (2024)
Scaling Group Inference for Diverse and High-Quality Generation
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
von: Parmar, Gaurav, et al.
Veröffentlicht: (2025)
Act2See: Emergent Active Visual Perception for Video Reasoning
von: Ma, Martin Q., et al.
Veröffentlicht: (2026)
von: Ma, Martin Q., et al.
Veröffentlicht: (2026)
Understanding Visual Concepts Across Models
von: Trabucco, Brandon, et al.
Veröffentlicht: (2024)
von: Trabucco, Brandon, et al.
Veröffentlicht: (2024)
Diffusion-DRF: Free, Rich, and Differentiable Reward for Video Diffusion Fine-Tuning
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
von: Wang, Yifan, et al.
Veröffentlicht: (2026)
Video Motion Transfer with Diffusion Transformers
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
von: Pondaven, Alexander, et al.
Veröffentlicht: (2024)
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2025)
Hierarchical Patch Diffusion Models for High-Resolution Video Generation
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2024)
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2024)
EFlow: Fast Few-Step Video Generator Training from Scratch via Efficient Solution Flow
von: Park, Dogyun, et al.
Veröffentlicht: (2026)
von: Park, Dogyun, et al.
Veröffentlicht: (2026)
Kontinuous Kontext: Continuous Strength Control for Instruction-based Image Editing
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
Can Text-to-Video Generation help Video-Language Alignment?
von: Zanella, Luca, et al.
Veröffentlicht: (2025)
von: Zanella, Luca, et al.
Veröffentlicht: (2025)
Video Active Perception: Effective Inference-Time Long-Form Video Understanding with Vision-Language Models
von: Ma, Martin Q., et al.
Veröffentlicht: (2026)
von: Ma, Martin Q., et al.
Veröffentlicht: (2026)
Effective Data Augmentation With Diffusion Models
von: Trabucco, Brandon, et al.
Veröffentlicht: (2023)
von: Trabucco, Brandon, et al.
Veröffentlicht: (2023)
Multi-subject Open-set Personalization in Video Generation
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
von: Chen, Tsai-Shien, et al.
Veröffentlicht: (2025)
Visual Concept-driven Image Generation with Text-to-Image Diffusion Model
von: Rahman, Tanzila, et al.
Veröffentlicht: (2024)
von: Rahman, Tanzila, et al.
Veröffentlicht: (2024)
AMAVA: Adaptive Motion-Aware Video-to-Audio Framework for Visually-Impaired Assistance
von: Klein, Benjamin, et al.
Veröffentlicht: (2026)
von: Klein, Benjamin, et al.
Veröffentlicht: (2026)
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
von: Dalal, Murtaza, et al.
Veröffentlicht: (2024)
Diffusion Priors for Dynamic View Synthesis from Monocular Videos
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
von: Wang, Chaoyang, et al.
Veröffentlicht: (2024)
Preventing Shortcuts in Adapter Training via Providing the Shortcuts
von: Goyal, Anujraaj Argo, et al.
Veröffentlicht: (2025)
von: Goyal, Anujraaj Argo, et al.
Veröffentlicht: (2025)
LayerComposer: Multi-Human Personalized Generation via Layered Canvas
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
von: Qian, Guocheng Gordon, et al.
Veröffentlicht: (2025)
Mind the Time: Temporally-Controlled Multi-Event Video Generation
von: Wu, Ziyi, et al.
Veröffentlicht: (2024)
von: Wu, Ziyi, et al.
Veröffentlicht: (2024)
JUST-DUB-IT: Video Dubbing via Joint Audio-Visual Diffusion
von: Chen, Anthony, et al.
Veröffentlicht: (2026)
von: Chen, Anthony, et al.
Veröffentlicht: (2026)
VIMI: Grounding Video Generation through Multi-modal Instruction
von: Fang, Yuwei, et al.
Veröffentlicht: (2024)
von: Fang, Yuwei, et al.
Veröffentlicht: (2024)
DenseDPO: Fine-Grained Temporal Preference Optimization for Video Diffusion Models
von: Wu, Ziyi, et al.
Veröffentlicht: (2025)
von: Wu, Ziyi, et al.
Veröffentlicht: (2025)
Towards Physical Understanding in Video Generation: A 3D Point Regularization Approach
von: Chen, Yunuo, et al.
Veröffentlicht: (2025)
von: Chen, Yunuo, et al.
Veröffentlicht: (2025)
In-Context Sync-LoRA for Portrait Video Editing
von: Polaczek, Sagi, et al.
Veröffentlicht: (2025)
von: Polaczek, Sagi, et al.
Veröffentlicht: (2025)
Consolidating Attention Features for Multi-view Image Editing
von: Patashnik, Or, et al.
Veröffentlicht: (2024)
von: Patashnik, Or, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Visual Personalization Turing Test
von: Abdal, Rameen, et al.
Veröffentlicht: (2026) -
Dynamic Concepts Personalization from Single Videos
von: Abdal, Rameen, et al.
Veröffentlicht: (2025) -
ArtifactLens: Hundreds of Labels Are Enough for Artifact Detection with VLMs
von: Burgess, James, et al.
Veröffentlicht: (2026) -
Zero-Shot Dynamic Concept Personalization with Grid-Based LoRA
von: Abdal, Rameen, et al.
Veröffentlicht: (2025) -
Omni-ID: Holistic Identity Representation Designed for Generative Tasks
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)