Generating Animated Layouts as Structured Text Representations
Fuente:
arXiv
Guardado en:
| Autores principales: | Shin, Yeonsang, Kim, Jihwan, Song, Yumin, Lee, Kyungseung, Chung, Hyunhee, Na, Taeyoung |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AGDC: Autoregressive Generation of Variable-Length Sequences with Joint Discrete and Continuous Spaces
por: Shin, Yeonsang, et al.
Publicado: (2026)
por: Shin, Yeonsang, et al.
Publicado: (2026)
TransText: Alpha-as-RGB Representation for Transparent Text Animation
por: Zhang, Fei, et al.
Publicado: (2026)
por: Zhang, Fei, et al.
Publicado: (2026)
ILDiff: Generate Transparent Animated Stickers by Implicit Layout Distillation
por: Zhang, Ting, et al.
Publicado: (2024)
por: Zhang, Ting, et al.
Publicado: (2024)
Learning to Sample Effective and Diverse Prompts for Text-to-Image Generation
por: Yun, Taeyoung, et al.
Publicado: (2025)
por: Yun, Taeyoung, et al.
Publicado: (2025)
Text-Animator: Controllable Visual Text Video Generation
por: Liu, Lin, et al.
Publicado: (2024)
por: Liu, Lin, et al.
Publicado: (2024)
Aggregated Structural Representation with Large Language Models for Human-Centric Layout Generation
por: Jin, Jiongchao, et al.
Publicado: (2025)
por: Jin, Jiongchao, et al.
Publicado: (2025)
SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models
por: Yang, Ruolin, et al.
Publicado: (2025)
por: Yang, Ruolin, et al.
Publicado: (2025)
Bidirectional Temporal Diffusion Model for Temporally Consistent Human Animation
por: Adiya, Tserendorj, et al.
Publicado: (2023)
por: Adiya, Tserendorj, et al.
Publicado: (2023)
FIFO-Diffusion: Generating Infinite Videos from Text without Training
por: Kim, Jihwan, et al.
Publicado: (2024)
por: Kim, Jihwan, et al.
Publicado: (2024)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
por: Park, NaHyeon, et al.
Publicado: (2024)
por: Park, NaHyeon, et al.
Publicado: (2024)
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking
por: Zheng, Zirui, et al.
Publicado: (2025)
por: Zheng, Zirui, et al.
Publicado: (2025)
Advancing Text-Driven Chest X-Ray Generation with Policy-Based Reinforcement Learning
por: Han, Woojung, et al.
Publicado: (2024)
por: Han, Woojung, et al.
Publicado: (2024)
StructLayoutFormer:Conditional Structured Layout Generation via Structure Serialization and Disentanglement
por: Hu, Xin, et al.
Publicado: (2025)
por: Hu, Xin, et al.
Publicado: (2025)
Activating Self-Attention for Multi-Scene Absolute Pose Regression
por: Lee, Miso, et al.
Publicado: (2024)
por: Lee, Miso, et al.
Publicado: (2024)
Long-term Pre-training for Temporal Action Detection with Transformers
por: Kim, Jihwan, et al.
Publicado: (2024)
por: Kim, Jihwan, et al.
Publicado: (2024)
Multitwine: Multi-Object Compositing with Text and Layout Control
por: Tarrés, Gemma Canet, et al.
Publicado: (2025)
por: Tarrés, Gemma Canet, et al.
Publicado: (2025)
Animate-X: Universal Character Image Animation with Enhanced Motion Representation
por: Tan, Shuai, et al.
Publicado: (2024)
por: Tan, Shuai, et al.
Publicado: (2024)
IM-Animation: An Implicit Motion Representation for Identity-decoupled Character Animation
por: Xu, Zhufeng, et al.
Publicado: (2026)
por: Xu, Zhufeng, et al.
Publicado: (2026)
Geometry-Aware Representation Denoising for Robust Multi-view 3D Reconstruction
por: Kim, Jin Hyeon, et al.
Publicado: (2026)
por: Kim, Jin Hyeon, et al.
Publicado: (2026)
TableSeq: Unified Generation of Structure, Content, and Layout
por: Hamdi, Laziz, et al.
Publicado: (2026)
por: Hamdi, Laziz, et al.
Publicado: (2026)
HAMLET: Switch your Vision-Language-Action Model into a History-Aware Policy
por: Koo, Myungkyu, et al.
Publicado: (2025)
por: Koo, Myungkyu, et al.
Publicado: (2025)
LayoutFlow: Flow Matching for Layout Generation
por: Guerreiro, Julian Jorge Andrade, et al.
Publicado: (2024)
por: Guerreiro, Julian Jorge Andrade, et al.
Publicado: (2024)
When Cars Have Stereotypes: Auditing Demographic Bias in Objects from Text-to-Image Models
por: Choi, Dasol, et al.
Publicado: (2025)
por: Choi, Dasol, et al.
Publicado: (2025)
FlipSketch: Flipping Static Drawings to Text-Guided Sketch Animations
por: Bandyopadhyay, Hmrishav, et al.
Publicado: (2024)
por: Bandyopadhyay, Hmrishav, et al.
Publicado: (2024)
SciPostLayout: A Dataset for Layout Analysis and Layout Generation of Scientific Posters
por: Tanaka, Shohei, et al.
Publicado: (2024)
por: Tanaka, Shohei, et al.
Publicado: (2024)
LED Benchmark: Diagnosing Structural Layout Errors for Document Layout Analysis
por: Heo, Inbum, et al.
Publicado: (2025)
por: Heo, Inbum, et al.
Publicado: (2025)
AnimateAnything: Consistent and Controllable Animation for Video Generation
por: Lei, Guojun, et al.
Publicado: (2024)
por: Lei, Guojun, et al.
Publicado: (2024)
Visual Diversity and Region-aware Prompt Learning for Zero-shot HOI Detection
por: Yang, Chanhyeong, et al.
Publicado: (2025)
por: Yang, Chanhyeong, et al.
Publicado: (2025)
GALA: Generating Animatable Layered Assets from a Single Scan
por: Kim, Taeksoo, et al.
Publicado: (2024)
por: Kim, Taeksoo, et al.
Publicado: (2024)
Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual Videos
por: Jeon, Subin, et al.
Publicado: (2024)
por: Jeon, Subin, et al.
Publicado: (2024)
TransAnimate: Taming Layer Diffusion to Generate RGBA Video
por: Chen, Xuewei, et al.
Publicado: (2025)
por: Chen, Xuewei, et al.
Publicado: (2025)
Retrieval-Augmented Layout Transformer for Content-Aware Layout Generation
por: Horita, Daichi, et al.
Publicado: (2023)
por: Horita, Daichi, et al.
Publicado: (2023)
LayoutDiffusion: Controllable Diffusion Model for Layout-to-image Generation
por: Zheng, Guangcong, et al.
Publicado: (2023)
por: Zheng, Guangcong, et al.
Publicado: (2023)
Visually Guided Generative Text-Layout Pre-training for Document Intelligence
por: Mao, Zhiming, et al.
Publicado: (2024)
por: Mao, Zhiming, et al.
Publicado: (2024)
DynamiCtrl: Rethinking the Basic Structure and the Role of Text for High-quality Human Image Animation
por: Zhao, Haoyu, et al.
Publicado: (2025)
por: Zhao, Haoyu, et al.
Publicado: (2025)
GaussianMotion: End-to-End Learning of Animatable Gaussian Avatars with Pose Guidance from Text
por: Shim, Gyumin, et al.
Publicado: (2025)
por: Shim, Gyumin, et al.
Publicado: (2025)
MORPHOS: Autoregressive 4D Generation with Temporal Structured Latents
por: Kwon, Minkyung, et al.
Publicado: (2026)
por: Kwon, Minkyung, et al.
Publicado: (2026)
Ctrl-Room: Controllable Text-to-3D Room Meshes Generation with Layout Constraints
por: Fang, Chuan, et al.
Publicado: (2023)
por: Fang, Chuan, et al.
Publicado: (2023)
Super-class guided Transformer for Zero-Shot Attribute Classification
por: Kim, Sehyung, et al.
Publicado: (2025)
por: Kim, Sehyung, et al.
Publicado: (2025)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
por: Cha, Junuk, et al.
Publicado: (2024)
por: Cha, Junuk, et al.
Publicado: (2024)
Ejemplares similares
-
AGDC: Autoregressive Generation of Variable-Length Sequences with Joint Discrete and Continuous Spaces
por: Shin, Yeonsang, et al.
Publicado: (2026) -
TransText: Alpha-as-RGB Representation for Transparent Text Animation
por: Zhang, Fei, et al.
Publicado: (2026) -
ILDiff: Generate Transparent Animated Stickers by Implicit Layout Distillation
por: Zhang, Ting, et al.
Publicado: (2024) -
Learning to Sample Effective and Diverse Prompts for Text-to-Image Generation
por: Yun, Taeyoung, et al.
Publicado: (2025) -
Text-Animator: Controllable Visual Text Video Generation
por: Liu, Lin, et al.
Publicado: (2024)