Character-Centered Dialogue Generation from Scene-Level Prompts
Fuente:
arXiv
Saved in:
| Main Authors: | Kang, Taewon, Lin, Ming C. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling
by: Kang, Taewon, et al.
Published: (2025)
by: Kang, Taewon, et al.
Published: (2025)
NEGATE: Constrained Semantic Guidance for Linguistic Negation in Text-to-Video Diffusion
by: Kang, Taewon, et al.
Published: (2026)
by: Kang, Taewon, et al.
Published: (2026)
Trajectory-Guided Diffusion for Foreground-Preserving Background Generation in Multi-Layer Documents
by: Kang, Taewon
Published: (2026)
by: Kang, Taewon
Published: (2026)
3D-free meets 3D priors: Novel View Synthesis from a Single Image with Pretrained Diffusion Guidance
by: Kang, Taewon, et al.
Published: (2024)
by: Kang, Taewon, et al.
Published: (2024)
DCR: Counterfactual Attractor Guidance for Rare Compositional Generation
by: Kang, Taewon, et al.
Published: (2026)
by: Kang, Taewon, et al.
Published: (2026)
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
by: Banerjee, Ayan, et al.
Published: (2025)
by: Banerjee, Ayan, et al.
Published: (2025)
Text-Conditioned Background Generation for Editable Multi-Layer Documents
by: Kang, Taewon, et al.
Published: (2025)
by: Kang, Taewon, et al.
Published: (2025)
HART: Human Aligned Reconstruction Transformer
by: Chen, Xiyi, et al.
Published: (2025)
by: Chen, Xiyi, et al.
Published: (2025)
Focus on the Whole Character: Discriminative Character Modeling for Scene Text Recognition
by: Zhou, Bangbang, et al.
Published: (2024)
by: Zhou, Bangbang, et al.
Published: (2024)
Char-SAM: Turning Segment Anything Model into Scene Text Segmentation Annotator with Character-level Visual Prompts
by: Xie, Enze, et al.
Published: (2024)
by: Xie, Enze, et al.
Published: (2024)
Single-View Scene Point Cloud Human Grasp Generation
by: Wang, Yan-Kang, et al.
Published: (2024)
by: Wang, Yan-Kang, et al.
Published: (2024)
Autonomous Character-Scene Interaction Synthesis from Text Instruction
by: Jiang, Nan, et al.
Published: (2024)
by: Jiang, Nan, et al.
Published: (2024)
LPA3D: 3D Room-Level Scene Generation from In-the-Wild Images
by: Yang, Ming-Jia, et al.
Published: (2025)
by: Yang, Ming-Jia, et al.
Published: (2025)
3D Scene Prompting for Scene-Consistent Camera-Controllable Video Generation
by: Lee, JoungBin, et al.
Published: (2025)
by: Lee, JoungBin, et al.
Published: (2025)
DynamicScaler: Seamless and Scalable Video Generation for Panoramic Scenes
by: Liu, Jinxiu, et al.
Published: (2024)
by: Liu, Jinxiu, et al.
Published: (2024)
VLPrompt: Vision-Language Prompting for Panoptic Scene Graph Generation
by: Zhou, Zijian, et al.
Published: (2023)
by: Zhou, Zijian, et al.
Published: (2023)
SceneLinker: Compositional 3D Scene Generation via Semantic Scene Graph from RGB Sequences
by: Kim, Seok-Young, et al.
Published: (2026)
by: Kim, Seok-Young, et al.
Published: (2026)
Make-A-Character 2: Animatable 3D Character Generation From a Single Image
by: Liu, Lin, et al.
Published: (2025)
by: Liu, Lin, et al.
Published: (2025)
The Fabrication of Reality and Fantasy: Scene Generation with LLM-Assisted Prompt Interpretation
by: Yao, Yi, et al.
Published: (2024)
by: Yao, Yi, et al.
Published: (2024)
Scene123: One Prompt to 3D Scene Generation via Video-Assisted and Consistency-Enhanced MAE
by: Yang, Yiying, et al.
Published: (2024)
by: Yang, Yiying, et al.
Published: (2024)
Character-Adapter: Prompt-Guided Region Control for High-Fidelity Character Customization
by: Ma, Yuhang, et al.
Published: (2024)
by: Ma, Yuhang, et al.
Published: (2024)
MegaScenes: Scene-Level View Synthesis at Scale
by: Tung, Joseph, et al.
Published: (2024)
by: Tung, Joseph, et al.
Published: (2024)
Unified Human-Scene Interaction via Prompted Chain-of-Contacts
by: Xiao, Zeqi, et al.
Published: (2023)
by: Xiao, Zeqi, et al.
Published: (2023)
Character Mixing for Video Generation
by: Liao, Tingting, et al.
Published: (2025)
by: Liao, Tingting, et al.
Published: (2025)
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder
by: Ma, Lichen, et al.
Published: (2024)
by: Ma, Lichen, et al.
Published: (2024)
Towards Lifelong Scene Graph Generation with Knowledge-ware In-context Prompt Learning
by: He, Tao, et al.
Published: (2024)
by: He, Tao, et al.
Published: (2024)
SceneDecorator: Towards Scene-Oriented Story Generation with Scene Planning and Scene Consistency
by: Song, Quanjian, et al.
Published: (2025)
by: Song, Quanjian, et al.
Published: (2025)
Fine-Grained Scene Graph Generation via Sample-Level Bias Prediction
by: Li, Yansheng, et al.
Published: (2024)
by: Li, Yansheng, et al.
Published: (2024)
Frame-Level Captions for Long Video Generation with Complex Multi Scenes
by: Zheng, Guangcong, et al.
Published: (2025)
by: Zheng, Guangcong, et al.
Published: (2025)
FreeScene: Mixed Graph Diffusion for 3D Scene Synthesis from Free Prompts
by: Bai, Tongyuan, et al.
Published: (2025)
by: Bai, Tongyuan, et al.
Published: (2025)
SceneGlue: Scene-Aware Transformer for Feature Matching without Scene-Level Annotation
by: Du, Songlin, et al.
Published: (2026)
by: Du, Songlin, et al.
Published: (2026)
PaintScene4D: Consistent 4D Scene Generation from Text Prompts
by: Gupta, Vinayak, et al.
Published: (2024)
by: Gupta, Vinayak, et al.
Published: (2024)
CharacterGen: Efficient 3D Character Generation from Single Images with Multi-View Pose Canonicalization
by: Peng, Hao-Yang, et al.
Published: (2024)
by: Peng, Hao-Yang, et al.
Published: (2024)
Controllable 3D Outdoor Scene Generation via Scene Graphs
by: Liu, Yuheng, et al.
Published: (2025)
by: Liu, Yuheng, et al.
Published: (2025)
SatDreamer360: Multiview-Consistent Generation of Ground-Level Scenes from Satellite Imagery
by: Ze, Xianghui, et al.
Published: (2025)
by: Ze, Xianghui, et al.
Published: (2025)
Contrastive Masked Autoencoders for Character-Level Open-Set Writer Identification
by: Jiang, Xiaowei, et al.
Published: (2025)
by: Jiang, Xiaowei, et al.
Published: (2025)
Chat-Edit-3D: Interactive 3D Scene Editing via Text Prompts
by: Fang, Shuangkang, et al.
Published: (2024)
by: Fang, Shuangkang, et al.
Published: (2024)
Generating Dialogues from Egocentric Instructional Videos for Task Assistance: Dataset, Method and Benchmark
by: Aggarwal, Lavisha, et al.
Published: (2025)
by: Aggarwal, Lavisha, et al.
Published: (2025)
Generating Multimodal Driving Scenes via Next-Scene Prediction
by: Wu, Yanhao, et al.
Published: (2025)
by: Wu, Yanhao, et al.
Published: (2025)
PromptRR: Diffusion Models as Prompt Generators for Single Image Reflection Removal
by: Wang, Tao, et al.
Published: (2024)
by: Wang, Tao, et al.
Published: (2024)
Similar Items
-
Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling
by: Kang, Taewon, et al.
Published: (2025) -
NEGATE: Constrained Semantic Guidance for Linguistic Negation in Text-to-Video Diffusion
by: Kang, Taewon, et al.
Published: (2026) -
Trajectory-Guided Diffusion for Foreground-Preserving Background Generation in Multi-Layer Documents
by: Kang, Taewon
Published: (2026) -
3D-free meets 3D priors: Novel View Synthesis from a Single Image with Pretrained Diffusion Guidance
by: Kang, Taewon, et al.
Published: (2024) -
DCR: Counterfactual Attractor Guidance for Rare Compositional Generation
by: Kang, Taewon, et al.
Published: (2026)