ILDiff: Generate Transparent Animated Stickers by Implicit Layout Distillation
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Ting, Yuan, Zhiqiang, Zhu, Yeshuang, Zhang, Jinchao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
WalkVLM:Aid Visually Impaired People Walking by Vision Language Model
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2024)
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2024)
RDTF: Resource-efficient Dual-mask Training Framework for Multi-frame Animated Sticker Generation
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2025)
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2025)
From Imitation to Innovation: The Emergence of AI Unique Artistic Styles and the Challenge of Copyright Protection
di: Jia, Zexi, et al.
Pubblicazione: (2025)
di: Jia, Zexi, et al.
Pubblicazione: (2025)
Animated Stickers: Bringing Stickers to Life with Video Diffusion
di: Yan, David, et al.
Pubblicazione: (2024)
di: Yan, David, et al.
Pubblicazione: (2024)
Control-CLIP: Decoupling Category and Style Guidance in CLIP for Specific-Domain Generation
di: Jia, Zexi, et al.
Pubblicazione: (2025)
di: Jia, Zexi, et al.
Pubblicazione: (2025)
VSD2M: A Large-scale Vision-language Sticker Dataset for Multi-frame Animated Sticker Generation
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2024)
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2024)
InterActHuman: Multi-Concept Human Animation with Layout-Aligned Audio Conditions
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
di: Wang, Zhenzhi, et al.
Pubblicazione: (2025)
Implicit Preference Alignment for Human Image Animation
di: Wang, Yuanzhi, et al.
Pubblicazione: (2026)
di: Wang, Yuanzhi, et al.
Pubblicazione: (2026)
Semantic to Structure: Learning Structural Representations for Infringement Detection
di: Huang, Chuanwei, et al.
Pubblicazione: (2025)
di: Huang, Chuanwei, et al.
Pubblicazione: (2025)
A Visual Leap in CLIP Compositionality Reasoning through Generation of Counterfactual Sets
di: Jia, Zexi, et al.
Pubblicazione: (2025)
di: Jia, Zexi, et al.
Pubblicazione: (2025)
AnimateDiff-Lightning: Cross-Model Diffusion Distillation
di: Lin, Shanchuan, et al.
Pubblicazione: (2024)
di: Lin, Shanchuan, et al.
Pubblicazione: (2024)
LoopAnimate: Loopable Salient Object Animation
di: Wang, Fanyi, et al.
Pubblicazione: (2024)
di: Wang, Fanyi, et al.
Pubblicazione: (2024)
Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation
di: Brioschi, Riccardo, et al.
Pubblicazione: (2025)
di: Brioschi, Riccardo, et al.
Pubblicazione: (2025)
Cafe-Talk: Generating 3D Talking Face Animation with Multimodal Coarse- and Fine-grained Control
di: Chen, Hejia, et al.
Pubblicazione: (2025)
di: Chen, Hejia, et al.
Pubblicazione: (2025)
Enhancing Visual Reliance in Text Generation: A Bayesian Perspective on Mitigating Hallucination in Large Vision-Language Models
di: Hu, Nanxing, et al.
Pubblicazione: (2025)
di: Hu, Nanxing, et al.
Pubblicazione: (2025)
PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction
di: Sun, Ting, et al.
Pubblicazione: (2025)
di: Sun, Ting, et al.
Pubblicazione: (2025)
MGHFT: Multi-Granularity Hierarchical Fusion Transformer for Cross-Modal Sticker Emotion Recognition
di: Chen, Jian, et al.
Pubblicazione: (2025)
di: Chen, Jian, et al.
Pubblicazione: (2025)
GSE: Evaluating Sticker Visual Semantic Similarity via a General Sticker Encoder
di: Chee, Heng Er Metilda, et al.
Pubblicazione: (2025)
di: Chee, Heng Er Metilda, et al.
Pubblicazione: (2025)
REFINE-CONTROL: A Semi-supervised Distillation Method For Conditional Image Generation
di: Jiang, Yicheng, et al.
Pubblicazione: (2025)
di: Jiang, Yicheng, et al.
Pubblicazione: (2025)
Can Large Models Fool the Eye? A New Turing Test for Biological Animation
di: Chen, Zijian, et al.
Pubblicazione: (2025)
di: Chen, Zijian, et al.
Pubblicazione: (2025)
StaR-KVQA: Structured Reasoning Traces for Implicit-Knowledge Visual Question Answering
di: Wen, Zhihao, et al.
Pubblicazione: (2025)
di: Wen, Zhihao, et al.
Pubblicazione: (2025)
AniGS: Animatable Gaussian Avatar from a Single Image with Inconsistent Gaussian Reconstruction
di: Qiu, Lingteng, et al.
Pubblicazione: (2024)
di: Qiu, Lingteng, et al.
Pubblicazione: (2024)
Animate Any Character in Any World
di: Wang, Yitong, et al.
Pubblicazione: (2025)
di: Wang, Yitong, et al.
Pubblicazione: (2025)
EvAnimate: Event-conditioned Image-to-Video Generation for Human Animation
di: Qu, Qiang, et al.
Pubblicazione: (2025)
di: Qu, Qiang, et al.
Pubblicazione: (2025)
DiffSpeaker: Speech-Driven 3D Facial Animation with Diffusion Transformer
di: Ma, Zhiyuan, et al.
Pubblicazione: (2024)
di: Ma, Zhiyuan, et al.
Pubblicazione: (2024)
PC-Talk: Precise Facial Animation Control for Audio-Driven Talking Face Generation
di: Wang, Baiqin, et al.
Pubblicazione: (2025)
di: Wang, Baiqin, et al.
Pubblicazione: (2025)
Order Is Not Layout: Order-to-Space Bias in Image Generation
di: Zhang, Yongkang, et al.
Pubblicazione: (2026)
di: Zhang, Yongkang, et al.
Pubblicazione: (2026)
Uneven Event Modeling for Partially Relevant Video Retrieval
di: Zhu, Sa, et al.
Pubblicazione: (2025)
di: Zhu, Sa, et al.
Pubblicazione: (2025)
Animating the Past: Reconstruct Trilobite via Video Generation
di: Wu, Xiaoran, et al.
Pubblicazione: (2024)
di: Wu, Xiaoran, et al.
Pubblicazione: (2024)
Global Intervention and Distillation for Federated Out-of-Distribution Generalization
di: Qi, Zhuang, et al.
Pubblicazione: (2025)
di: Qi, Zhuang, et al.
Pubblicazione: (2025)
Zero-shot High-fidelity and Pose-controllable Character Animation
di: Zhu, Bingwen, et al.
Pubblicazione: (2024)
di: Zhu, Bingwen, et al.
Pubblicazione: (2024)
DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance
di: Luo, Yuxuan, et al.
Pubblicazione: (2025)
di: Luo, Yuxuan, et al.
Pubblicazione: (2025)
GeoDM: Geometry-aware Distribution Matching for Dataset Distillation
di: Li, Xuhui, et al.
Pubblicazione: (2025)
di: Li, Xuhui, et al.
Pubblicazione: (2025)
OmniAlpha: Aligning Transparency-Aware Generation via Multi-Task Unified Reinforcement Learning
di: Yu, Hao, et al.
Pubblicazione: (2025)
di: Yu, Hao, et al.
Pubblicazione: (2025)
See Before You Code: Learning Visual Priors for Spatially Aware Educational Animation Generation
di: Li, Yuejia, et al.
Pubblicazione: (2026)
di: Li, Yuejia, et al.
Pubblicazione: (2026)
Training-free Composite Scene Generation for Layout-to-Image Synthesis
di: Liu, Jiaqi, et al.
Pubblicazione: (2024)
di: Liu, Jiaqi, et al.
Pubblicazione: (2024)
SceneLLM: Implicit Language Reasoning in LLM for Dynamic Scene Graph Generation
di: Zhang, Hang, et al.
Pubblicazione: (2024)
di: Zhang, Hang, et al.
Pubblicazione: (2024)
Dataset Distillation via Committee Voting
di: Cui, Jiacheng, et al.
Pubblicazione: (2025)
di: Cui, Jiacheng, et al.
Pubblicazione: (2025)
StableAnimator: High-Quality Identity-Preserving Human Image Animation
di: Tu, Shuyuan, et al.
Pubblicazione: (2024)
di: Tu, Shuyuan, et al.
Pubblicazione: (2024)
Automatic Layout Planning for Visually-Rich Documents with Instruction-Following Models
di: Zhu, Wanrong, et al.
Pubblicazione: (2024)
di: Zhu, Wanrong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
WalkVLM:Aid Visually Impaired People Walking by Vision Language Model
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2024) -
RDTF: Resource-efficient Dual-mask Training Framework for Multi-frame Animated Sticker Generation
di: Yuan, Zhiqiang, et al.
Pubblicazione: (2025) -
From Imitation to Innovation: The Emergence of AI Unique Artistic Styles and the Challenge of Copyright Protection
di: Jia, Zexi, et al.
Pubblicazione: (2025) -
Animated Stickers: Bringing Stickers to Life with Video Diffusion
di: Yan, David, et al.
Pubblicazione: (2024) -
Control-CLIP: Decoupling Category and Style Guidance in CLIP for Specific-Domain Generation
di: Jia, Zexi, et al.
Pubblicazione: (2025)