Evolving Storytelling: Benchmarks and Methods for New Character Customization with Diffusion Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Xiyu, Wang, Yufei, Tsutsui, Satoshi, Lin, Weisi, Wen, Bihan, Kot, Alex C. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Robust and Reliable Concept Representations: Reliability-Enhanced Concept Embedding Model
di: Cai, Yuxuan, et al.
Pubblicazione: (2025)
di: Cai, Yuxuan, et al.
Pubblicazione: (2025)
DP-IQA: Utilizing Diffusion Prior for Blind Image Quality Assessment in the Wild
di: Fu, Honghao, et al.
Pubblicazione: (2024)
di: Fu, Honghao, et al.
Pubblicazione: (2024)
From Chaos to Clarity: 3DGS in the Dark
di: Li, Zhihao, et al.
Pubblicazione: (2024)
di: Li, Zhihao, et al.
Pubblicazione: (2024)
Integrating Clinical Knowledge into Concept Bottleneck Models
di: Pang, Winnie, et al.
Pubblicazione: (2024)
di: Pang, Winnie, et al.
Pubblicazione: (2024)
Digital Staining with Knowledge Distillation: A Unified Framework for Unpaired and Paired-But-Misaligned Data
di: Xu, Ziwang, et al.
Pubblicazione: (2025)
di: Xu, Ziwang, et al.
Pubblicazione: (2025)
WBCAtt+: Fine-Grained Pixel-Level Morphological Annotations for White Blood Cell Images
di: Tsutsui, Satoshi, et al.
Pubblicazione: (2026)
di: Tsutsui, Satoshi, et al.
Pubblicazione: (2026)
ContextGS: Compact 3D Gaussian Splatting with Anchor Level Context Model
di: Wang, Yufei, et al.
Pubblicazione: (2024)
di: Wang, Yufei, et al.
Pubblicazione: (2024)
Discovering Hidden Visual Concepts Beyond Linguistic Input in Infant Learning
di: Ke, Xueyi, et al.
Pubblicazione: (2025)
di: Ke, Xueyi, et al.
Pubblicazione: (2025)
Reconciling Stochastic and Deterministic Strategies for Zero-shot Image Restoration using Diffusion Model in Dual
di: Wang, Chong, et al.
Pubblicazione: (2025)
di: Wang, Chong, et al.
Pubblicazione: (2025)
Single-Image Shadow Removal Using Deep Learning: A Comprehensive Survey
di: Guo, Laniqng, et al.
Pubblicazione: (2024)
di: Guo, Laniqng, et al.
Pubblicazione: (2024)
SoftShadow: Leveraging Soft Masks for Penumbra-Aware Shadow Removal
di: Wang, Xinrui, et al.
Pubblicazione: (2024)
di: Wang, Xinrui, et al.
Pubblicazione: (2024)
Sparc3D: Sparse Representation and Construction for High-Resolution 3D Shapes Modeling
di: Li, Zhihao, et al.
Pubblicazione: (2025)
di: Li, Zhihao, et al.
Pubblicazione: (2025)
Benchmarking Adversarial Robustness of Image Shadow Removal with Shadow-adaptive Attacks
di: Wang, Chong, et al.
Pubblicazione: (2024)
di: Wang, Chong, et al.
Pubblicazione: (2024)
RealisticDreamer: Guidance Score Distillation for Few-shot Gaussian Splatting
di: Wu, Ruocheng, et al.
Pubblicazione: (2025)
di: Wu, Ruocheng, et al.
Pubblicazione: (2025)
Customized Visual Storytelling with Unified Multimodal LLMs
di: Li, Wei-Hua, et al.
Pubblicazione: (2026)
di: Li, Wei-Hua, et al.
Pubblicazione: (2026)
Skeleton-to-Image Encoding: Enabling Skeleton Representation Learning via Vision-Pretrained Models
di: Yang, Siyuan, et al.
Pubblicazione: (2026)
di: Yang, Siyuan, et al.
Pubblicazione: (2026)
CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models
di: Wang, Qinghe, et al.
Pubblicazione: (2024)
di: Wang, Qinghe, et al.
Pubblicazione: (2024)
Image Quality Assessment for Machines: Paradigm, Large-scale Database, and Models
di: Wang, Xiaoqi, et al.
Pubblicazione: (2025)
di: Wang, Xiaoqi, et al.
Pubblicazione: (2025)
InstantCharacter: Personalize Any Characters with a Scalable Diffusion Transformer Framework
di: Tao, Jiale, et al.
Pubblicazione: (2025)
di: Tao, Jiale, et al.
Pubblicazione: (2025)
MUGSQA: Novel Multi-Uncertainty-Based Gaussian Splatting Quality Assessment Method, Dataset, and Benchmarks
di: Chen, Tianang, et al.
Pubblicazione: (2025)
di: Chen, Tianang, et al.
Pubblicazione: (2025)
SAVER: Mitigating Hallucinations in Large Vision-Language Models via Style-Aware Visual Early Revision
di: Li, Zhaoxu, et al.
Pubblicazione: (2025)
di: Li, Zhaoxu, et al.
Pubblicazione: (2025)
Persistent Story World Simulation with Continuous Character Customization
di: Zhang, Jinlu, et al.
Pubblicazione: (2026)
di: Zhang, Jinlu, et al.
Pubblicazione: (2026)
DreamVTON: Customizing 3D Virtual Try-on with Personalized Diffusion Models
di: Xie, Zhenyu, et al.
Pubblicazione: (2024)
di: Xie, Zhenyu, et al.
Pubblicazione: (2024)
Training-Free Text-Guided Image Editing with Visual Autoregressive Model
di: Wang, Yufei, et al.
Pubblicazione: (2025)
di: Wang, Yufei, et al.
Pubblicazione: (2025)
Make a Cheap Scaling: A Self-Cascade Diffusion Model for Higher-Resolution Adaptation
di: Guo, Lanqing, et al.
Pubblicazione: (2024)
di: Guo, Lanqing, et al.
Pubblicazione: (2024)
BADiff: Bandwidth Adaptive Diffusion Model
di: Zhang, Xi, et al.
Pubblicazione: (2025)
di: Zhang, Xi, et al.
Pubblicazione: (2025)
Color Space Learning for Cross-Color Person Re-Identification
di: Nie, Jiahao, et al.
Pubblicazione: (2024)
di: Nie, Jiahao, et al.
Pubblicazione: (2024)
See What You Seek: Semantic Contextual Integration for Cloth-Changing Person Re-Identification
di: Han, Xiyu, et al.
Pubblicazione: (2024)
di: Han, Xiyu, et al.
Pubblicazione: (2024)
Enhancing Diffusion Models with Text-Encoder Reinforcement Learning
di: Chen, Chaofeng, et al.
Pubblicazione: (2023)
di: Chen, Chaofeng, et al.
Pubblicazione: (2023)
SimBase: A Simple Baseline for Temporal Video Grounding
di: Bao, Peijun, et al.
Pubblicazione: (2024)
di: Bao, Peijun, et al.
Pubblicazione: (2024)
PIDiff: Image Customization for Personalized Identities with Diffusion Models
di: Gu, Jinyu, et al.
Pubblicazione: (2025)
di: Gu, Jinyu, et al.
Pubblicazione: (2025)
Intelligent Grimm -- Open-ended Visual Storytelling via Latent Diffusion Models
di: Liu, Chang, et al.
Pubblicazione: (2023)
di: Liu, Chang, et al.
Pubblicazione: (2023)
ActivityForensics: A Comprehensive Benchmark for Localizing Manipulated Activity in Videos
di: Bao, Peijun, et al.
Pubblicazione: (2026)
di: Bao, Peijun, et al.
Pubblicazione: (2026)
MMRel: Benchmarking Relation Understanding in Multi-Modal Large Language Models
di: Nie, Jiahao, et al.
Pubblicazione: (2024)
di: Nie, Jiahao, et al.
Pubblicazione: (2024)
Non-confusing Generation of Customized Concepts in Diffusion Models
di: Lin, Wang, et al.
Pubblicazione: (2024)
di: Lin, Wang, et al.
Pubblicazione: (2024)
Q-Bench+: A Benchmark for Multi-modal Foundation Models on Low-level Vision from Single Images to Pairs
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
di: Zhang, Zicheng, et al.
Pubblicazione: (2024)
IdentiFace: Multi-Modal Iterative Diffusion Framework for Identifiable Suspect Face Generation in Crime Investigations
di: Liu, Weichen, et al.
Pubblicazione: (2026)
di: Liu, Weichen, et al.
Pubblicazione: (2026)
RF4D:Neural Radar Fields for Novel View Synthesis in Outdoor Dynamic Scenes
di: Zhang, Jiarui, et al.
Pubblicazione: (2025)
di: Zhang, Jiarui, et al.
Pubblicazione: (2025)
Progressive Divide-and-Conquer via Subsampling Decomposition for Accelerated MRI
di: Wang, Chong, et al.
Pubblicazione: (2024)
di: Wang, Chong, et al.
Pubblicazione: (2024)
Temporal As a Plugin: Unsupervised Video Denoising with Pre-Trained Image Denoisers
di: Fu, Zixuan, et al.
Pubblicazione: (2024)
di: Fu, Zixuan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Towards Robust and Reliable Concept Representations: Reliability-Enhanced Concept Embedding Model
di: Cai, Yuxuan, et al.
Pubblicazione: (2025) -
DP-IQA: Utilizing Diffusion Prior for Blind Image Quality Assessment in the Wild
di: Fu, Honghao, et al.
Pubblicazione: (2024) -
From Chaos to Clarity: 3DGS in the Dark
di: Li, Zhihao, et al.
Pubblicazione: (2024) -
Integrating Clinical Knowledge into Concept Bottleneck Models
di: Pang, Winnie, et al.
Pubblicazione: (2024) -
Digital Staining with Knowledge Distillation: A Unified Framework for Unpaired and Paired-But-Misaligned Data
di: Xu, Ziwang, et al.
Pubblicazione: (2025)