OneActor: Consistent Character Generation via Cluster-Conditioned Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Jiahao, Yan, Caixia, Lin, Haonan, Zhang, Weizhan, Wang, Mengmeng, Gong, Tieliang, Dai, Guang, Sun, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpotActor: Training-Free Layout-Controlled Consistent Image Generation
by: Wang, Jiahao, et al.
Published: (2024)
by: Wang, Jiahao, et al.
Published: (2024)
MVGSR: Multi-View Consistent 3D Gaussian Super-Resolution via Epipolar Guidance
by: Zhang, Kaizhe, et al.
Published: (2025)
by: Zhang, Kaizhe, et al.
Published: (2025)
Accelerating Non-Maximum Suppression: A Graph Theory Perspective
by: Si, King-Siong, et al.
Published: (2024)
by: Si, King-Siong, et al.
Published: (2024)
Rectified Diffusion Guidance for Conditional Generation
by: Xia, Mengfei, et al.
Published: (2024)
by: Xia, Mengfei, et al.
Published: (2024)
Flipped Classroom: Aligning Teacher Attention with Student in Generalized Category Discovery
by: Lin, Haonan, et al.
Published: (2024)
by: Lin, Haonan, et al.
Published: (2024)
EchoShot: Multi-Shot Portrait Video Generation
by: Wang, Jiahao, et al.
Published: (2025)
by: Wang, Jiahao, et al.
Published: (2025)
AnyID: Ultra-Fidelity Universal Identity-Preserving Video Generation from Any Visual References
by: Wang, Jiahao, et al.
Published: (2026)
by: Wang, Jiahao, et al.
Published: (2026)
InfoSAM: Fine-Tuning the Segment Anything Model from An Information-Theoretic Perspective
by: Zhang, Yuanhong, et al.
Published: (2025)
by: Zhang, Yuanhong, et al.
Published: (2025)
Schedule Your Edit: A Simple yet Effective Diffusion Noise Schedule for Image Editing
by: Lin, Haonan, et al.
Published: (2024)
by: Lin, Haonan, et al.
Published: (2024)
DreamSalon: A Staged Diffusion Framework for Preserving Identity-Context in Editable Face Generation
by: Lin, Haonan, et al.
Published: (2024)
by: Lin, Haonan, et al.
Published: (2024)
Towards the Generalization of Multi-view Learning: An Information-theoretical Analysis
by: Wen, Wen, et al.
Published: (2025)
by: Wen, Wen, et al.
Published: (2025)
Consistent123: One Image to Highly Consistent 3D Asset Using Case-Aware Diffusion Priors
by: Lin, Yukang, et al.
Published: (2023)
by: Lin, Yukang, et al.
Published: (2023)
CharaConsist: Fine-Grained Consistent Character Generation
by: Wang, Mengyu, et al.
Published: (2025)
by: Wang, Mengyu, et al.
Published: (2025)
Tile Classification Based Viewport Prediction with Multi-modal Fusion Transformer
by: Zhang, Zhihao, et al.
Published: (2023)
by: Zhang, Zhihao, et al.
Published: (2023)
Unsupervised Structural Scene Decomposition via Foreground-Aware Slot Attention with Pseudo-Mask Guidance
by: Sheng, Huankun, et al.
Published: (2025)
by: Sheng, Huankun, et al.
Published: (2025)
DepthArb: Training-Free Depth-Arbitrated Generation for Occlusion-Robust Image Synthesis
by: Niu, Hongjin, et al.
Published: (2026)
by: Niu, Hongjin, et al.
Published: (2026)
Exploring Time Conditioning in Diffusion Generative Models from Disjoint Noisy Data Manifolds
by: Li, Liuzhuozheng, et al.
Published: (2026)
by: Li, Liuzhuozheng, et al.
Published: (2026)
AnyCharV: Bootstrap Controllable Character Video Generation with Fine-to-Coarse Guidance
by: Wang, Zhao, et al.
Published: (2025)
by: Wang, Zhao, et al.
Published: (2025)
BachVid: Training-Free Video Generation with Consistent Background and Character
by: Yan, Han, et al.
Published: (2025)
by: Yan, Han, et al.
Published: (2025)
No Other Representation Component Is Needed: Diffusion Transformers Can Provide Representation Guidance by Themselves
by: Jiang, Dengyang, et al.
Published: (2025)
by: Jiang, Dengyang, et al.
Published: (2025)
How Does Distribution Matching Help Domain Generalization: An Information-theoretic Analysis
by: Dong, Yuxin, et al.
Published: (2024)
by: Dong, Yuxin, et al.
Published: (2024)
DynamicID: Zero-Shot Multi-ID Image Personalization with Flexible Facial Editability
by: Hu, Xirui, et al.
Published: (2025)
by: Hu, Xirui, et al.
Published: (2025)
Make-A-Character 2: Animatable 3D Character Generation From a Single Image
by: Liu, Lin, et al.
Published: (2025)
by: Liu, Lin, et al.
Published: (2025)
Gloria: Consistent Character Video Generation via Content Anchors
by: Yang, Yuhang, et al.
Published: (2026)
by: Yang, Yuhang, et al.
Published: (2026)
Consistent Video Colorization via Palette Guidance
by: Wang, Han, et al.
Published: (2025)
by: Wang, Han, et al.
Published: (2025)
Conditional Text-to-Image Generation with Reference Guidance
by: Kim, Taewook, et al.
Published: (2024)
by: Kim, Taewook, et al.
Published: (2024)
Information-Theoretic Generalization Bounds of Replay-based Continual Learning
by: Wen, Wen, et al.
Published: (2025)
by: Wen, Wen, et al.
Published: (2025)
CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models
by: Wang, Qinghe, et al.
Published: (2024)
by: Wang, Qinghe, et al.
Published: (2024)
CharacterShot: Controllable and Consistent 4D Character Animation
by: Gao, Junyao, et al.
Published: (2025)
by: Gao, Junyao, et al.
Published: (2025)
Self-Adversarial One Step Generation via Condition Shifting
by: Liu, Deyuan, et al.
Published: (2026)
by: Liu, Deyuan, et al.
Published: (2026)
The Chosen One: Consistent Characters in Text-to-Image Diffusion Models
by: Avrahami, Omri, et al.
Published: (2023)
by: Avrahami, Omri, et al.
Published: (2023)
Low-Biased General Annotated Dataset Generation
by: Jiang, Dengyang, et al.
Published: (2024)
by: Jiang, Dengyang, et al.
Published: (2024)
MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model
by: Dai, Wenxun, et al.
Published: (2024)
by: Dai, Wenxun, et al.
Published: (2024)
Conditional Polarization Guidance for Camouflaged Object Detection
by: Zhang, QIfan, et al.
Published: (2026)
by: Zhang, QIfan, et al.
Published: (2026)
Follow-Your-MultiPose: Tuning-Free Multi-Character Text-to-Video Generation via Pose Guidance
by: Zhang, Beiyuan, et al.
Published: (2024)
by: Zhang, Beiyuan, et al.
Published: (2024)
HAODiff: Human-Aware One-Step Diffusion via Dual-Prompt Guidance
by: Gong, Jue, et al.
Published: (2025)
by: Gong, Jue, et al.
Published: (2025)
Deep Incomplete Multi-view Clustering with Distribution Dual-Consistency Recovery Guidance
by: Jin, Jiaqi, et al.
Published: (2025)
by: Jin, Jiaqi, et al.
Published: (2025)
Instructing Text-to-Image Diffusion Models via Classifier-Guided Semantic Optimization
by: Chang, Yuanyuan, et al.
Published: (2025)
by: Chang, Yuanyuan, et al.
Published: (2025)
FancyVideo: Towards Dynamic and Consistent Video Generation via Cross-frame Textual Guidance
by: Feng, Jiasong, et al.
Published: (2024)
by: Feng, Jiasong, et al.
Published: (2024)
ReMix: Towards a Unified View of Consistent Character Generation and Editing
by: Zhou, Benjia, et al.
Published: (2025)
by: Zhou, Benjia, et al.
Published: (2025)
Similar Items
-
SpotActor: Training-Free Layout-Controlled Consistent Image Generation
by: Wang, Jiahao, et al.
Published: (2024) -
MVGSR: Multi-View Consistent 3D Gaussian Super-Resolution via Epipolar Guidance
by: Zhang, Kaizhe, et al.
Published: (2025) -
Accelerating Non-Maximum Suppression: A Graph Theory Perspective
by: Si, King-Siong, et al.
Published: (2024) -
Rectified Diffusion Guidance for Conditional Generation
by: Xia, Mengfei, et al.
Published: (2024) -
Flipped Classroom: Aligning Teacher Attention with Student in Generalized Category Discovery
by: Lin, Haonan, et al.
Published: (2024)