Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Hwang, Sungwon, Jang, Hyojin, Kim, Kinam, Park, Minho, Choo, Jaegul |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
por: Kim, Kinam, et al.
Publicado: (2025)
por: Kim, Kinam, et al.
Publicado: (2025)
SphereDiff: Tuning-free 360° Static and Dynamic Panorama Generation via Spherical Latent Representation
por: Park, Minho, et al.
Publicado: (2025)
por: Park, Minho, et al.
Publicado: (2025)
EgoX: Egocentric Video Generation from a Single Exocentric Video
por: Kang, Taewoong, et al.
Publicado: (2025)
por: Kang, Taewoong, et al.
Publicado: (2025)
InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion
por: Jin, Hoiyeong, et al.
Publicado: (2025)
por: Jin, Hoiyeong, et al.
Publicado: (2025)
Spatiotemporal Skip Guidance for Enhanced Video Diffusion Sampling
por: Hyung, Junha, et al.
Publicado: (2024)
por: Hyung, Junha, et al.
Publicado: (2024)
Good Noise Makes Good Edits: A Training-Free Diffusion-Based Video Editing with Image and Text Prompts
por: Choi, Saemee, et al.
Publicado: (2025)
por: Choi, Saemee, et al.
Publicado: (2025)
CA-LoRA: Concept-Aware LoRA for Domain-Aligned Segmentation Dataset Generation
por: Park, Minho, et al.
Publicado: (2025)
por: Park, Minho, et al.
Publicado: (2025)
Regularized Training with Generated Datasets for Name-Only Transfer of Vision-Language Models
por: Park, Minho, et al.
Publicado: (2024)
por: Park, Minho, et al.
Publicado: (2024)
Zero-Shot Head Swapping in Real-World Scenarios
por: Kang, Taewoong, et al.
Publicado: (2025)
por: Kang, Taewoong, et al.
Publicado: (2025)
VEGS: View Extrapolation of Urban Scenes in 3D Gaussian Splatting using Learned Priors
por: Hwang, Sungwon, et al.
Publicado: (2024)
por: Hwang, Sungwon, et al.
Publicado: (2024)
SurFhead: Affine Rig Blending for Geometrically Accurate 2D Gaussian Surfel Head Avatars
por: Lee, Jaeseong, et al.
Publicado: (2024)
por: Lee, Jaeseong, et al.
Publicado: (2024)
Effective Rank Analysis and Regularization for Enhanced 3D Gaussian Splatting
por: Hyung, Junha, et al.
Publicado: (2024)
por: Hyung, Junha, et al.
Publicado: (2024)
What to Preserve and What to Transfer: Faithful, Identity-Preserving Diffusion-based Hairstyle Transfer
por: Chung, Chaeyeon, et al.
Publicado: (2024)
por: Chung, Chaeyeon, et al.
Publicado: (2024)
Memory-Efficient Fine-Tuning Diffusion Transformers via Dynamic Patch Sampling and Block Skipping
por: Park, Sunghyun, et al.
Publicado: (2026)
por: Park, Sunghyun, et al.
Publicado: (2026)
Devil is in the Detail: Towards Injecting Fine Details of Image Prompt in Image Generation via Conflict-free Guidance and Stratified Attention
por: Jo, Kyungmin, et al.
Publicado: (2025)
por: Jo, Kyungmin, et al.
Publicado: (2025)
TV-LiVE: Training-Free, Text-Guided Video Editing via Layer Informed Vitality Exploitation
por: Kim, Min-Jung, et al.
Publicado: (2025)
por: Kim, Min-Jung, et al.
Publicado: (2025)
DiffuseSlide: Training-Free High Frame Rate Video Generation Diffusion
por: Hwang, Geunmin, et al.
Publicado: (2025)
por: Hwang, Geunmin, et al.
Publicado: (2025)
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models
por: Kim, Jeongho, et al.
Publicado: (2024)
por: Kim, Jeongho, et al.
Publicado: (2024)
Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure
por: Yun, Jooyeol, et al.
Publicado: (2025)
por: Yun, Jooyeol, et al.
Publicado: (2025)
Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects
por: Jo, Kyungmin, et al.
Publicado: (2024)
por: Jo, Kyungmin, et al.
Publicado: (2024)
Scaling Up Personalized Image Aesthetic Assessment via Task Vector Customization
por: Yun, Jooyeol, et al.
Publicado: (2024)
por: Yun, Jooyeol, et al.
Publicado: (2024)
Frame Guidance: Training-Free Guidance for Frame-Level Control in Video Diffusion Models
por: Jang, Sangwon, et al.
Publicado: (2025)
por: Jang, Sangwon, et al.
Publicado: (2025)
Frame-wise Conditioning Adaptation for Fine-Tuning Diffusion Models in Text-to-Video Prediction
por: Liu, Zheyuan, et al.
Publicado: (2025)
por: Liu, Zheyuan, et al.
Publicado: (2025)
Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation
por: Kim, Min-Jung, et al.
Publicado: (2025)
por: Kim, Min-Jung, et al.
Publicado: (2025)
Towards Calibrated Robust Fine-Tuning of Vision-Language Models
por: Oh, Changdae, et al.
Publicado: (2023)
por: Oh, Changdae, et al.
Publicado: (2023)
SHIFT: Motion Alignment in Video Diffusion Models with Adversarial Hybrid Fine-Tuning
por: Ye, Xi, et al.
Publicado: (2026)
por: Ye, Xi, et al.
Publicado: (2026)
Bones Can't Be Triangles: Accurate and Efficient Vertebrae Keypoint Estimation through Collaborative Error Revision
por: Kim, Jinhee, et al.
Publicado: (2024)
por: Kim, Jinhee, et al.
Publicado: (2024)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
por: Kim, Jeongho, et al.
Publicado: (2024)
por: Kim, Jeongho, et al.
Publicado: (2024)
Enhancing Intrinsic Features for Debiasing via Investigating Class-Discerning Common Attributes in Bias-Contrastive Pair
por: Park, Jeonghoon, et al.
Publicado: (2024)
por: Park, Jeonghoon, et al.
Publicado: (2024)
From Wardrobe to Canvas: Wardrobe Polyptych LoRA for Part-level Controllable Human Image Generation
por: Kim, Jeongho, et al.
Publicado: (2025)
por: Kim, Jeongho, et al.
Publicado: (2025)
Training Spatial-Frequency Visual Prompts and Probabilistic Clusters for Accurate Black-Box Transfer Learning
por: Cho, Wonwoo, et al.
Publicado: (2024)
por: Cho, Wonwoo, et al.
Publicado: (2024)
AHS: Adaptive Head Synthesis via Synthetic Data Augmentations
por: Kang, Taewoong, et al.
Publicado: (2026)
por: Kang, Taewoong, et al.
Publicado: (2026)
GaussianMotion: End-to-End Learning of Animatable Gaussian Avatars with Pose Guidance from Text
por: Shim, Gyumin, et al.
Publicado: (2025)
por: Shim, Gyumin, et al.
Publicado: (2025)
Generalizable Disaster Damage Assessment via Change Detection with Vision Foundation Model
por: Ahn, Kyeongjin, et al.
Publicado: (2024)
por: Ahn, Kyeongjin, et al.
Publicado: (2024)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
por: Hwang, Dongyoon, et al.
Publicado: (2024)
por: Hwang, Dongyoon, et al.
Publicado: (2024)
Event-Based Video Frame Interpolation With Cross-Modal Asymmetric Bidirectional Motion Fields
por: Kim, Taewoo, et al.
Publicado: (2025)
por: Kim, Taewoo, et al.
Publicado: (2025)
Enabling Region-Specific Control via Lassos in Point-Based Colorization
por: Lee, Sanghyeon, et al.
Publicado: (2024)
por: Lee, Sanghyeon, et al.
Publicado: (2024)
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
por: Kim, Donghu, et al.
Publicado: (2024)
por: Kim, Donghu, et al.
Publicado: (2024)
Advancing Cross-Domain Generalizability in Face Anti-Spoofing: Insights, Design, and Metrics
por: Kim, Hyojin, et al.
Publicado: (2024)
por: Kim, Hyojin, et al.
Publicado: (2024)
Enhanced OoD Detection through Cross-Modal Alignment of Multi-Modal Representations
por: Kim, Jeonghyeon, et al.
Publicado: (2025)
por: Kim, Jeonghyeon, et al.
Publicado: (2025)
Ejemplares similares
-
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
por: Kim, Kinam, et al.
Publicado: (2025) -
SphereDiff: Tuning-free 360° Static and Dynamic Panorama Generation via Spherical Latent Representation
por: Park, Minho, et al.
Publicado: (2025) -
EgoX: Egocentric Video Generation from a Single Exocentric Video
por: Kang, Taewoong, et al.
Publicado: (2025) -
InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion
por: Jin, Hoiyeong, et al.
Publicado: (2025) -
Spatiotemporal Skip Guidance for Enhanced Video Diffusion Sampling
por: Hyung, Junha, et al.
Publicado: (2024)