2K-Characters-10K-Stories: A Quality-Gated Stylized Narrative Dataset with Disentangled Control and Sequence Consistency
Fuente:
arXiv
Saved in:
| Main Authors: | Yin, Xingxi, Li, Yicheng, Yan, Gong, Li, Chenglin, Zhao, Jian, Huang, Cong, Deng, Yue, Zhang, Yin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
InstructAttribute: Fine-grained Object Attributes editing with Instruction
by: Yin, Xingxi, et al.
Published: (2025)
by: Yin, Xingxi, et al.
Published: (2025)
ColorEdit: Training-free Image-Guided Color editing with diffusion model
by: Yin, Xingxi, et al.
Published: (2024)
by: Yin, Xingxi, et al.
Published: (2024)
VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning
by: Li, Chenglin, et al.
Published: (2026)
by: Li, Chenglin, et al.
Published: (2026)
RelationAdapter: Learning and Transferring Visual Relation with Diffusion Transformers
by: Gong, Yan, et al.
Published: (2025)
by: Gong, Yan, et al.
Published: (2025)
StoryMaker: Towards Holistic Consistent Characters in Text-to-image Generation
by: Zhou, Zhengguang, et al.
Published: (2024)
by: Zhou, Zhengguang, et al.
Published: (2024)
TheaterGen: Character Management with LLM for Consistent Multi-turn Image Generation
by: Cheng, Junhao, et al.
Published: (2024)
by: Cheng, Junhao, et al.
Published: (2024)
K12Vista: Exploring the Boundaries of MLLMs in K-12 Education
by: Li, Chong, et al.
Published: (2025)
by: Li, Chong, et al.
Published: (2025)
Gate-Based and Annealing-Based Quantum Algorithms for the Maximum K-Plex Problem
by: Li, Xiaofan, et al.
Published: (2025)
by: Li, Xiaofan, et al.
Published: (2025)
GaussianBlender: Instant Stylization of 3D Gaussians with Disentangled Latent Spaces
by: Ocal, Melis, et al.
Published: (2025)
by: Ocal, Melis, et al.
Published: (2025)
Sequence-to-Sequence Language Models for Character and Emotion Detection in Dream Narratives
by: Cortal, Gustave
Published: (2024)
by: Cortal, Gustave
Published: (2024)
ReDiStory: Region-Disentangled Diffusion for Consistent Visual Story Generation
by: Sarkar, Ayushman, et al.
Published: (2026)
by: Sarkar, Ayushman, et al.
Published: (2026)
CharacterShot: Controllable and Consistent 4D Character Animation
by: Gao, Junyao, et al.
Published: (2025)
by: Gao, Junyao, et al.
Published: (2025)
DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion
by: He, Huiguo, et al.
Published: (2024)
by: He, Huiguo, et al.
Published: (2024)
Locations of Characters in Narratives: Andersen and Persuasion Datasets
by: Ozyurt, Batuhan, et al.
Published: (2025)
by: Ozyurt, Batuhan, et al.
Published: (2025)
S2ED: From Story to Executable Descriptions for Consistency-Aware Story Illustration
by: Yin, Sijing, et al.
Published: (2026)
by: Yin, Sijing, et al.
Published: (2026)
The GPT-WritingPrompts Dataset: A Comparative Analysis of Character Portrayal in Short Stories
by: Huang, Xi Yu, et al.
Published: (2024)
by: Huang, Xi Yu, et al.
Published: (2024)
Free-Lunch Color-Texture Disentanglement for Stylized Image Generation
by: Qin, Jiang, et al.
Published: (2025)
by: Qin, Jiang, et al.
Published: (2025)
CHATTER: A Character Attribution Dataset for Narrative Understanding
by: Baruah, Sabyasachee, et al.
Published: (2024)
by: Baruah, Sabyasachee, et al.
Published: (2024)
Multilingual Synopses of Movie Narratives: A Dataset for Vision-Language Story Understanding
by: Sun, Yidan, et al.
Published: (2024)
by: Sun, Yidan, et al.
Published: (2024)
Mixed Distillation Helps Smaller Language Model Better Reasoning
by: Li, Chenglin, et al.
Published: (2023)
by: Li, Chenglin, et al.
Published: (2023)
Advancing Marine Research: UWSAM Framework and UIIS10K Dataset for Precise Underwater Instance Segmentation
by: Li, Hua, et al.
Published: (2025)
by: Li, Hua, et al.
Published: (2025)
STORYANCHORS: Generating Consistent Multi-Scene Story Frames for Long-Form Narratives
by: Wang, Bo, et al.
Published: (2025)
by: Wang, Bo, et al.
Published: (2025)
InfinityStory: Unlimited Video Generation with World Consistency and Character-Aware Shot Transitions
by: Elmoghany, Mohamed, et al.
Published: (2026)
by: Elmoghany, Mohamed, et al.
Published: (2026)
Synchronized Multi‐Frame Diffusion for Temporally Consistent Video Stylization
by: Minshan Xie, et al.
Published: (2025)
by: Minshan Xie, et al.
Published: (2025)
VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models
by: Li, Chenglin, et al.
Published: (2024)
by: Li, Chenglin, et al.
Published: (2024)
DEADiff: An Efficient Stylization Diffusion Model with Disentangled Representations
by: Qi, Tianhao, et al.
Published: (2024)
by: Qi, Tianhao, et al.
Published: (2024)
Balancing Stylization and Truth via Disentangled Representation Steering
by: Shen, Chenglei, et al.
Published: (2025)
by: Shen, Chenglei, et al.
Published: (2025)
Realizing Degree Sequences With S3‐Connected Graphs
by: Rui Guan, et al.
Published: (2026)
by: Rui Guan, et al.
Published: (2026)
CharacterFactory: Sampling Consistent Characters with GANs for Diffusion Models
by: Wang, Qinghe, et al.
Published: (2024)
by: Wang, Qinghe, et al.
Published: (2024)
Lights, Camera, Consistency: A Multistage Pipeline for Character-Stable AI Video Stories
by: Jain, Chayan, et al.
Published: (2025)
by: Jain, Chayan, et al.
Published: (2025)
A search for $B_{s0}^{*}$ and $B^{*}_{s1}$ through the $K^{-}p$ interaction
by: Yuan, Min, et al.
Published: (2025)
by: Yuan, Min, et al.
Published: (2025)
USIS16K: High-Quality Dataset for Underwater Salient Instance Segmentation
by: Hong, Lin, et al.
Published: (2025)
by: Hong, Lin, et al.
Published: (2025)
MVQA-68K: A Multi-dimensional and Causally-annotated Dataset with Quality Interpretability for Video Assessment
by: Pu, Yanyun, et al.
Published: (2025)
by: Pu, Yanyun, et al.
Published: (2025)
Optimizing Instruction Synthesis: Effective Exploration of Evolutionary Space with Tree Search
by: Li, Chenglin, et al.
Published: (2024)
by: Li, Chenglin, et al.
Published: (2024)
Lost in Stories: Consistency Bugs in Long Story Generation by LLMs
by: Li, Junjie, et al.
Published: (2026)
by: Li, Junjie, et al.
Published: (2026)
Iterative Zoom-In: Temporal Interval Exploration for Long Video Understanding
by: Li, Chenglin, et al.
Published: (2025)
by: Li, Chenglin, et al.
Published: (2025)
MegaHan97K: A Large-Scale Dataset for Mega-Category Chinese Character Recognition with over 97K Categories
by: Zhang, Yuyi, et al.
Published: (2025)
by: Zhang, Yuyi, et al.
Published: (2025)
RACon: Retrieval-Augmented Simulated Character Locomotion Control
by: Mu, Yuxuan, et al.
Published: (2024)
by: Mu, Yuxuan, et al.
Published: (2024)
Air Quality Prediction with A Meteorology-Guided Modality-Decoupled Spatio-Temporal Network
by: Yin, Hang, et al.
Published: (2025)
by: Yin, Hang, et al.
Published: (2025)
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
by: Hu, Li, et al.
Published: (2023)
by: Hu, Li, et al.
Published: (2023)
Similar Items
-
InstructAttribute: Fine-grained Object Attributes editing with Instruction
by: Yin, Xingxi, et al.
Published: (2025) -
ColorEdit: Training-free Image-Guided Color editing with diffusion model
by: Yin, Xingxi, et al.
Published: (2024) -
VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning
by: Li, Chenglin, et al.
Published: (2026) -
RelationAdapter: Learning and Transferring Visual Relation with Diffusion Transformers
by: Gong, Yan, et al.
Published: (2025) -
StoryMaker: Towards Holistic Consistent Characters in Text-to-image Generation
by: Zhou, Zhengguang, et al.
Published: (2024)