Evolving Storytelling: Benchmarks and Methods for New Character Customization with Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Xiyu, Wang, Yufei, Tsutsui, Satoshi, Lin, Weisi, Wen, Bihan, Kot, Alex C. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Robust and Reliable Concept Representations: Reliability-Enhanced Concept Embedding Model
by: Cai, Yuxuan, et al.
Published: (2025)
by: Cai, Yuxuan, et al.
Published: (2025)
DP-IQA: Utilizing Diffusion Prior for Blind Image Quality Assessment in the Wild
by: Fu, Honghao, et al.
Published: (2024)
by: Fu, Honghao, et al.
Published: (2024)
From Chaos to Clarity: 3DGS in the Dark
by: Li, Zhihao, et al.
Published: (2024)
by: Li, Zhihao, et al.
Published: (2024)
Integrating Clinical Knowledge into Concept Bottleneck Models
by: Pang, Winnie, et al.
Published: (2024)
by: Pang, Winnie, et al.
Published: (2024)
Digital Staining with Knowledge Distillation: A Unified Framework for Unpaired and Paired-But-Misaligned Data
by: Xu, Ziwang, et al.
Published: (2025)
by: Xu, Ziwang, et al.
Published: (2025)
WBCAtt+: Fine-Grained Pixel-Level Morphological Annotations for White Blood Cell Images
by: Tsutsui, Satoshi, et al.
Published: (2026)
by: Tsutsui, Satoshi, et al.
Published: (2026)
ContextGS: Compact 3D Gaussian Splatting with Anchor Level Context Model
by: Wang, Yufei, et al.
Published: (2024)
by: Wang, Yufei, et al.
Published: (2024)
Discovering Hidden Visual Concepts Beyond Linguistic Input in Infant Learning
by: Ke, Xueyi, et al.
Published: (2025)
by: Ke, Xueyi, et al.
Published: (2025)
Reconciling Stochastic and Deterministic Strategies for Zero-shot Image Restoration using Diffusion Model in Dual
by: Wang, Chong, et al.
Published: (2025)
by: Wang, Chong, et al.
Published: (2025)
Single-Image Shadow Removal Using Deep Learning: A Comprehensive Survey
by: Guo, Laniqng, et al.
Published: (2024)
by: Guo, Laniqng, et al.
Published: (2024)
SoftShadow: Leveraging Soft Masks for Penumbra-Aware Shadow Removal
by: Wang, Xinrui, et al.
Published: (2024)
by: Wang, Xinrui, et al.
Published: (2024)
Sparc3D: Sparse Representation and Construction for High-Resolution 3D Shapes Modeling
by: Li, Zhihao, et al.
Published: (2025)
by: Li, Zhihao, et al.
Published: (2025)
Benchmarking Adversarial Robustness of Image Shadow Removal with Shadow-adaptive Attacks
by: Wang, Chong, et al.
Published: (2024)
by: Wang, Chong, et al.
Published: (2024)
RealisticDreamer: Guidance Score Distillation for Few-shot Gaussian Splatting
by: Wu, Ruocheng, et al.
Published: (2025)
by: Wu, Ruocheng, et al.
Published: (2025)
Customized Visual Storytelling with Unified Multimodal LLMs
by: Li, Wei-Hua, et al.
Published: (2026)
by: Li, Wei-Hua, et al.
Published: (2026)
Skeleton-to-Image Encoding: Enabling Skeleton Representation Learning via Vision-Pretrained Models
by: Yang, Siyuan, et al.
Published: (2026)
by: Yang, Siyuan, et al.
Published: (2026)
CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models
by: Wang, Qinghe, et al.
Published: (2024)
by: Wang, Qinghe, et al.
Published: (2024)
Image Quality Assessment for Machines: Paradigm, Large-scale Database, and Models
by: Wang, Xiaoqi, et al.
Published: (2025)
by: Wang, Xiaoqi, et al.
Published: (2025)
InstantCharacter: Personalize Any Characters with a Scalable Diffusion Transformer Framework
by: Tao, Jiale, et al.
Published: (2025)
by: Tao, Jiale, et al.
Published: (2025)
MUGSQA: Novel Multi-Uncertainty-Based Gaussian Splatting Quality Assessment Method, Dataset, and Benchmarks
by: Chen, Tianang, et al.
Published: (2025)
by: Chen, Tianang, et al.
Published: (2025)
SAVER: Mitigating Hallucinations in Large Vision-Language Models via Style-Aware Visual Early Revision
by: Li, Zhaoxu, et al.
Published: (2025)
by: Li, Zhaoxu, et al.
Published: (2025)
Persistent Story World Simulation with Continuous Character Customization
by: Zhang, Jinlu, et al.
Published: (2026)
by: Zhang, Jinlu, et al.
Published: (2026)
DreamVTON: Customizing 3D Virtual Try-on with Personalized Diffusion Models
by: Xie, Zhenyu, et al.
Published: (2024)
by: Xie, Zhenyu, et al.
Published: (2024)
Training-Free Text-Guided Image Editing with Visual Autoregressive Model
by: Wang, Yufei, et al.
Published: (2025)
by: Wang, Yufei, et al.
Published: (2025)
Make a Cheap Scaling: A Self-Cascade Diffusion Model for Higher-Resolution Adaptation
by: Guo, Lanqing, et al.
Published: (2024)
by: Guo, Lanqing, et al.
Published: (2024)
BADiff: Bandwidth Adaptive Diffusion Model
by: Zhang, Xi, et al.
Published: (2025)
by: Zhang, Xi, et al.
Published: (2025)
Color Space Learning for Cross-Color Person Re-Identification
by: Nie, Jiahao, et al.
Published: (2024)
by: Nie, Jiahao, et al.
Published: (2024)
See What You Seek: Semantic Contextual Integration for Cloth-Changing Person Re-Identification
by: Han, Xiyu, et al.
Published: (2024)
by: Han, Xiyu, et al.
Published: (2024)
Enhancing Diffusion Models with Text-Encoder Reinforcement Learning
by: Chen, Chaofeng, et al.
Published: (2023)
by: Chen, Chaofeng, et al.
Published: (2023)
SimBase: A Simple Baseline for Temporal Video Grounding
by: Bao, Peijun, et al.
Published: (2024)
by: Bao, Peijun, et al.
Published: (2024)
PIDiff: Image Customization for Personalized Identities with Diffusion Models
by: Gu, Jinyu, et al.
Published: (2025)
by: Gu, Jinyu, et al.
Published: (2025)
Intelligent Grimm -- Open-ended Visual Storytelling via Latent Diffusion Models
by: Liu, Chang, et al.
Published: (2023)
by: Liu, Chang, et al.
Published: (2023)
ActivityForensics: A Comprehensive Benchmark for Localizing Manipulated Activity in Videos
by: Bao, Peijun, et al.
Published: (2026)
by: Bao, Peijun, et al.
Published: (2026)
MMRel: Benchmarking Relation Understanding in Multi-Modal Large Language Models
by: Nie, Jiahao, et al.
Published: (2024)
by: Nie, Jiahao, et al.
Published: (2024)
Non-confusing Generation of Customized Concepts in Diffusion Models
by: Lin, Wang, et al.
Published: (2024)
by: Lin, Wang, et al.
Published: (2024)
Q-Bench+: A Benchmark for Multi-modal Foundation Models on Low-level Vision from Single Images to Pairs
by: Zhang, Zicheng, et al.
Published: (2024)
by: Zhang, Zicheng, et al.
Published: (2024)
IdentiFace: Multi-Modal Iterative Diffusion Framework for Identifiable Suspect Face Generation in Crime Investigations
by: Liu, Weichen, et al.
Published: (2026)
by: Liu, Weichen, et al.
Published: (2026)
RF4D:Neural Radar Fields for Novel View Synthesis in Outdoor Dynamic Scenes
by: Zhang, Jiarui, et al.
Published: (2025)
by: Zhang, Jiarui, et al.
Published: (2025)
Progressive Divide-and-Conquer via Subsampling Decomposition for Accelerated MRI
by: Wang, Chong, et al.
Published: (2024)
by: Wang, Chong, et al.
Published: (2024)
Temporal As a Plugin: Unsupervised Video Denoising with Pre-Trained Image Denoisers
by: Fu, Zixuan, et al.
Published: (2024)
by: Fu, Zixuan, et al.
Published: (2024)
Similar Items
-
Towards Robust and Reliable Concept Representations: Reliability-Enhanced Concept Embedding Model
by: Cai, Yuxuan, et al.
Published: (2025) -
DP-IQA: Utilizing Diffusion Prior for Blind Image Quality Assessment in the Wild
by: Fu, Honghao, et al.
Published: (2024) -
From Chaos to Clarity: 3DGS in the Dark
by: Li, Zhihao, et al.
Published: (2024) -
Integrating Clinical Knowledge into Concept Bottleneck Models
by: Pang, Winnie, et al.
Published: (2024) -
Digital Staining with Knowledge Distillation: A Unified Framework for Unpaired and Paired-But-Misaligned Data
by: Xu, Ziwang, et al.
Published: (2025)