Identity-Preserving Image-to-Video Generation via Reward-Guided Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Liao, Jiang, Wentao, Zhu, Yiran, Li, Jiahe, Ge, Tiezheng, Cao, Zhiguo, Zheng, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AtomoVideo: High Fidelity Image-to-Video Generation
by: Gong, Litong, et al.
Published: (2024)
by: Gong, Litong, et al.
Published: (2024)
Tuning-Free Noise Rectification for High Fidelity Image-to-Video Generation
by: Li, Weijie, et al.
Published: (2024)
by: Li, Weijie, et al.
Published: (2024)
Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing
by: Xu, Shaodong, et al.
Published: (2026)
by: Xu, Shaodong, et al.
Published: (2026)
VC4VG: Optimizing Video Captions for Text-to-Video Generation
by: Du, Yang, et al.
Published: (2025)
by: Du, Yang, et al.
Published: (2025)
RISE-T2V: Rephrasing and Injecting Semantics with LLM for Expansive Text-to-Video Generation
by: Zhang, Xiangjun, et al.
Published: (2025)
by: Zhang, Xiangjun, et al.
Published: (2025)
Rethinking Scribble-Guided Image Editing: Generalization, Instruction Adherence, and Multi-Tasking
by: Xu, Mingyi, et al.
Published: (2026)
by: Xu, Mingyi, et al.
Published: (2026)
DMM: Building a Versatile Image Generation Model via Distillation-Based Model Merging
by: Song, Tianhui, et al.
Published: (2025)
by: Song, Tianhui, et al.
Published: (2025)
Accelerating Image Generation with Sub-path Linear Approximation Model
by: Xu, Chen, et al.
Published: (2024)
by: Xu, Chen, et al.
Published: (2024)
WildActor: Unconstrained Identity-Preserving Video Generation
by: Guo, Qin, et al.
Published: (2026)
by: Guo, Qin, et al.
Published: (2026)
Identity-GRPO: Optimizing Multi-Human Identity-preserving Video Generation via Reinforcement Learning
by: Meng, Xiangyu, et al.
Published: (2025)
by: Meng, Xiangyu, et al.
Published: (2025)
Identity-Preserving Text-to-Video Generation by Frequency Decomposition
by: Yuan, Shenghai, et al.
Published: (2024)
by: Yuan, Shenghai, et al.
Published: (2024)
FlowDCN: Exploring DCN-like Architectures for Fast Image Generation with Arbitrary Resolution
by: Wang, Shuai, et al.
Published: (2024)
by: Wang, Shuai, et al.
Published: (2024)
End-to-end Video Gaze Estimation via Capturing Head-face-eye Spatial-temporal Interaction Context
by: Guan, Yiran, et al.
Published: (2023)
by: Guan, Yiran, et al.
Published: (2023)
Identity-Preserving Text-to-Video Generation Guided by Simple yet Effective Spatial-Temporal Decoupled Representations
by: Wang, Yuji, et al.
Published: (2025)
by: Wang, Yuji, et al.
Published: (2025)
Decoupled Competitive Framework for Semi-supervised Medical Image Segmentation
by: Chen, Jiahe, et al.
Published: (2025)
by: Chen, Jiahe, et al.
Published: (2025)
Towards Robust Monocular Depth Estimation in Non-Lambertian Surfaces
by: Zhang, Junrui, et al.
Published: (2024)
by: Zhang, Junrui, et al.
Published: (2024)
Identity-Preserving Text-to-Image Generation via Dual-Level Feature Decoupling and Expert-Guided Fusion
by: Chen, Kewen, et al.
Published: (2025)
by: Chen, Kewen, et al.
Published: (2025)
AdvDMD: Adversarial Reward Meets DMD For High-Quality Few-Step Generation
by: Wang, Xu, et al.
Published: (2026)
by: Wang, Xu, et al.
Published: (2026)
ConsID-Gen: View-Consistent and Identity-Preserving Image-to-Video Generation
by: Wu, Mingyang, et al.
Published: (2026)
by: Wu, Mingyang, et al.
Published: (2026)
ID-Animator: Zero-Shot Identity-Preserving Human Video Generation
by: He, Xuanhua, et al.
Published: (2024)
by: He, Xuanhua, et al.
Published: (2024)
Enhancing Prompt Following with Visual Control Through Training-Free Mask-Guided Diffusion
by: Chen, Hongyu, et al.
Published: (2024)
by: Chen, Hongyu, et al.
Published: (2024)
T-Stars-Poster: A Framework for Product-Centric Advertising Image Design
by: Chen, Hongyu, et al.
Published: (2025)
by: Chen, Hongyu, et al.
Published: (2025)
Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement
by: Gao, Jiayi, et al.
Published: (2025)
by: Gao, Jiayi, et al.
Published: (2025)
Show and Polish: Reference-Guided Identity Preservation in Face Video Restoration
by: Han, Wenkang, et al.
Published: (2025)
by: Han, Wenkang, et al.
Published: (2025)
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback Learning
by: Chen, Weifeng, et al.
Published: (2024)
by: Chen, Weifeng, et al.
Published: (2024)
Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation
by: Lu, Yunhong, et al.
Published: (2025)
by: Lu, Yunhong, et al.
Published: (2025)
RHanDS: Refining Malformed Hands for Generated Images with Decoupled Structure and Style Guidance
by: Wang, Chengrui, et al.
Published: (2024)
by: Wang, Chengrui, et al.
Published: (2024)
DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation
by: Chen, Junhao, et al.
Published: (2025)
by: Chen, Junhao, et al.
Published: (2025)
DOLLAR: Few-Step Video Generation via Distillation and Latent Reward Optimization
by: Ding, Zihan, et al.
Published: (2024)
by: Ding, Zihan, et al.
Published: (2024)
MoCA: Identity-Preserving Text-to-Video Generation via Mixture of Cross Attention
by: Xie, Qi, et al.
Published: (2025)
by: Xie, Qi, et al.
Published: (2025)
Lifelong Histopathology Whole Slide Image Retrieval via Distance Consistency Rehearsal
by: Zhu, Xinyu, et al.
Published: (2024)
by: Zhu, Xinyu, et al.
Published: (2024)
EchoVideo: Identity-Preserving Human Video Generation by Multimodal Feature Fusion
by: Wei, Jiangchuan, et al.
Published: (2025)
by: Wei, Jiangchuan, et al.
Published: (2025)
RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards
by: Ye, Junyan, et al.
Published: (2025)
by: Ye, Junyan, et al.
Published: (2025)
PosterMaker: Towards High-Quality Product Poster Generation with Accurate Text Rendering
by: Gao, Yifan, et al.
Published: (2025)
by: Gao, Yifan, et al.
Published: (2025)
Identity-Preserving Video Dubbing Using Motion Warping
by: Liu, Runzhen, et al.
Published: (2025)
by: Liu, Runzhen, et al.
Published: (2025)
DyBluRF: Dynamic Neural Radiance Fields from Blurry Monocular Video
by: Sun, Huiqiang, et al.
Published: (2024)
by: Sun, Huiqiang, et al.
Published: (2024)
Facial Demorphing via Identity Preserving Image Decomposition
by: Shukla, Nitish, et al.
Published: (2024)
by: Shukla, Nitish, et al.
Published: (2024)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
by: Lai, Yixuan, et al.
Published: (2026)
by: Lai, Yixuan, et al.
Published: (2026)
Identity-Preserving Pose-Guided Character Animation via Facial Landmarks Transformation
by: Mu, Lianrui, et al.
Published: (2024)
by: Mu, Lianrui, et al.
Published: (2024)
Geo-Align: Video Generation Alignment via Metric Geometry Reward
by: Li, Zizun, et al.
Published: (2026)
by: Li, Zizun, et al.
Published: (2026)
Similar Items
-
AtomoVideo: High Fidelity Image-to-Video Generation
by: Gong, Litong, et al.
Published: (2024) -
Tuning-Free Noise Rectification for High Fidelity Image-to-Video Generation
by: Li, Weijie, et al.
Published: (2024) -
Edit-GRPO: A Locality-Preserving Policy Optimization Framework for Image Editing
by: Xu, Shaodong, et al.
Published: (2026) -
VC4VG: Optimizing Video Captions for Text-to-Video Generation
by: Du, Yang, et al.
Published: (2025) -
RISE-T2V: Rephrasing and Injecting Semantics with LLM for Expansive Text-to-Video Generation
by: Zhang, Xiangjun, et al.
Published: (2025)