Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models
Fuente:
arXiv
Saved in:
| Main Authors: | Jeon, Wooseok, Park, Seungho, Shin, Seunghyun, Lee, Sangeyl, Jeong, Hyeonho, Jeon, Hae-Gon |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Motion Prior Distillation in Time Reversal Sampling for Generative Inbetweening
by: Jeon, Wooseok, et al.
Published: (2026)
by: Jeon, Wooseok, et al.
Published: (2026)
Scene Structure Guidance Network: Unfolding Graph Partitioning into Pixel-Wise Feature Learning
by: Shin, Jisu, et al.
Published: (2023)
by: Shin, Jisu, et al.
Published: (2023)
Kinetic Typography Diffusion Model
by: Park, Seonmi, et al.
Published: (2024)
by: Park, Seonmi, et al.
Published: (2024)
Video Color Grading via Look-Up Table Generation
by: Shin, Seunghyun, et al.
Published: (2025)
by: Shin, Seunghyun, et al.
Published: (2025)
Inverse Image-Based Rendering for Light Field Generation from Single Images
by: Jung, Hyunjun, et al.
Published: (2025)
by: Jung, Hyunjun, et al.
Published: (2025)
Spectral Motion Alignment for Video Motion Transfer using Diffusion Models
by: Park, Geon Yeong, et al.
Published: (2024)
by: Park, Geon Yeong, et al.
Published: (2024)
Universal Image Immunization against Diffusion-based Image Editing via Semantic Injection
by: Lee, Chanhui, et al.
Published: (2026)
by: Lee, Chanhui, et al.
Published: (2026)
CHROMA: Consistent Harmonization of Multi-View Appearance via Bilateral Grid Prediction
by: Shin, Jisu, et al.
Published: (2025)
by: Shin, Jisu, et al.
Published: (2025)
Reangle-A-Video: 4D Video Generation as Video-to-Video Translation
by: Jeong, Hyeonho, et al.
Published: (2025)
by: Jeong, Hyeonho, et al.
Published: (2025)
Depth Prompting for Sensor-Agnostic Depth Estimation
by: Park, Jin-Hwi, et al.
Published: (2024)
by: Park, Jin-Hwi, et al.
Published: (2024)
DreamMotion: Space-Time Self-Similar Score Distillation for Zero-Shot Video Editing
by: Jeong, Hyeonho, et al.
Published: (2024)
by: Jeong, Hyeonho, et al.
Published: (2024)
V.I.P. : Iterative Online Preference Distillation for Efficient Video Diffusion Models
by: Kim, Jisoo, et al.
Published: (2025)
by: Kim, Jisoo, et al.
Published: (2025)
SingularTrajectory: Universal Trajectory Predictor Using Diffusion Model
by: Bae, Inhwan, et al.
Published: (2024)
by: Bae, Inhwan, et al.
Published: (2024)
Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models
by: Jeong, Hyeonho, et al.
Published: (2023)
by: Jeong, Hyeonho, et al.
Published: (2023)
Continuous Locomotive Crowd Behavior Generation
by: Bae, Inhwan, et al.
Published: (2025)
by: Bae, Inhwan, et al.
Published: (2025)
Improving Video Diffusion Transformer Training by Multi-Feature Fusion and Alignment from Self-Supervised Vision Encoders
by: Lee, Dohun, et al.
Published: (2025)
by: Lee, Dohun, et al.
Published: (2025)
CanonicalFusion: Generating Drivable 3D Human Avatars from Multiple Images
by: Shin, Jisu, et al.
Published: (2024)
by: Shin, Jisu, et al.
Published: (2024)
Can Language Beat Numerical Regression? Language-Based Multimodal Trajectory Prediction
by: Bae, Inhwan, et al.
Published: (2024)
by: Bae, Inhwan, et al.
Published: (2024)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
by: Jeong, Hyeonho, et al.
Published: (2024)
by: Jeong, Hyeonho, et al.
Published: (2024)
JOG3R: Towards 3D-Consistent Video Generators
by: Huang, Chun-Hao Paul, et al.
Published: (2025)
by: Huang, Chun-Hao Paul, et al.
Published: (2025)
SPG: Improving Motion Diffusion by Smooth Perturbation Guidance
by: Jeon, Boseong
Published: (2025)
by: Jeon, Boseong
Published: (2025)
RehearsalNeRF: Decoupling Intrinsic Neural Fields of Dynamic Illuminations for Scene Editing
by: Won, Changyeon, et al.
Published: (2026)
by: Won, Changyeon, et al.
Published: (2026)
CollabVR: Collaborative Video Reasoning with Vision-Language and Video Generation Models
by: Kim, Joowon, et al.
Published: (2026)
by: Kim, Joowon, et al.
Published: (2026)
ComPose: When to Trust Hands for Object Pose Tracking
by: Shin, Jisu, et al.
Published: (2026)
by: Shin, Jisu, et al.
Published: (2026)
Relaxed Rigidity with Ray-based Grouping for Dynamic Gaussian Splatting
by: Lee, Junoh, et al.
Published: (2026)
by: Lee, Junoh, et al.
Published: (2026)
Fully Explicit Dynamic Gaussian Splatting
by: Lee, Junoh, et al.
Published: (2024)
by: Lee, Junoh, et al.
Published: (2024)
Data-driven Precipitation Nowcasting Using Satellite Imagery
by: Park, Young-Jae, et al.
Published: (2024)
by: Park, Young-Jae, et al.
Published: (2024)
Disentangled Motion Modeling for Video Frame Interpolation
by: Lew, Jaihyun, et al.
Published: (2024)
by: Lew, Jaihyun, et al.
Published: (2024)
FrameBridge: Improving Image-to-Video Generation with Bridge Models
by: Wang, Yuji, et al.
Published: (2024)
by: Wang, Yuji, et al.
Published: (2024)
EM-Vid: Training-Free Entity-Centric Memory for Efficient and Consistent Multi-Shot Video Generation
by: Vandersanden, Jente, et al.
Published: (2026)
by: Vandersanden, Jente, et al.
Published: (2026)
Improving Motion in Image-to-Video Models via Adaptive Low-Pass Guidance
by: Choi, June Suk, et al.
Published: (2025)
by: Choi, June Suk, et al.
Published: (2025)
Leveraging Out-of-Distribution Unlabeled Images: Semi-Supervised Semantic Segmentation with an Open-Vocabulary Model
by: Shin, Wooseok, et al.
Published: (2025)
by: Shin, Wooseok, et al.
Published: (2025)
AgentRVOS: Reasoning over Object Tracks for Zero-Shot Referring Video Object Segmentation
by: Jin, Woojeong, et al.
Published: (2026)
by: Jin, Woojeong, et al.
Published: (2026)
Memory-V2V: Memory-Augmented Video-to-Video Diffusion for Consistent Multi-Turn Editing
by: Lee, Dohun, et al.
Published: (2026)
by: Lee, Dohun, et al.
Published: (2026)
MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation
by: Lee, Minhyun, et al.
Published: (2024)
by: Lee, Minhyun, et al.
Published: (2024)
Vision-aligned Latent Reasoning for Multi-modal Large Language Model
by: Jeon, Byungwoo, et al.
Published: (2026)
by: Jeon, Byungwoo, et al.
Published: (2026)
PAC-FNO: Parallel-Structured All-Component Fourier Neural Operators for Recognizing Low-Quality Images
by: Jeon, Jinsung, et al.
Published: (2024)
by: Jeon, Jinsung, et al.
Published: (2024)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
by: Kim, Kangyeol, et al.
Published: (2024)
by: Kim, Kangyeol, et al.
Published: (2024)
Investigation of Frame Differences as Motion Cues for Video Object Segmentation
by: Kawamura, Sota, et al.
Published: (2025)
by: Kawamura, Sota, et al.
Published: (2025)
Video Finetuning Improves Reasoning Between Frames
by: Yang, Ruiqi, et al.
Published: (2025)
by: Yang, Ruiqi, et al.
Published: (2025)
Similar Items
-
Motion Prior Distillation in Time Reversal Sampling for Generative Inbetweening
by: Jeon, Wooseok, et al.
Published: (2026) -
Scene Structure Guidance Network: Unfolding Graph Partitioning into Pixel-Wise Feature Learning
by: Shin, Jisu, et al.
Published: (2023) -
Kinetic Typography Diffusion Model
by: Park, Seonmi, et al.
Published: (2024) -
Video Color Grading via Look-Up Table Generation
by: Shin, Seunghyun, et al.
Published: (2025) -
Inverse Image-Based Rendering for Light Field Generation from Single Images
by: Jung, Hyunjun, et al.
Published: (2025)