Narrative-to-Scene Generation: An LLM-Driven Pipeline for 2D Game Environments
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Yi-Chun, Jhala, Arnav |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GameTileNet: A Semantic Dataset for Low-Resolution Game Art in Procedural Content Generation
di: Chen, Yi-Chun, et al.
Pubblicazione: (2025)
di: Chen, Yi-Chun, et al.
Pubblicazione: (2025)
Think2Sing: Orchestrating Structured Motion Subtitles for Singing-Driven 3D Head Animation
di: Huang, Zikai, et al.
Pubblicazione: (2025)
di: Huang, Zikai, et al.
Pubblicazione: (2025)
PediatricsMQA: a Multi-modal Pediatrics Question Answering Benchmark
di: Bahaj, Adil, et al.
Pubblicazione: (2025)
di: Bahaj, Adil, et al.
Pubblicazione: (2025)
Multimodal Cinematic Video Synthesis Using Text-to-Image and Audio Generation Models
di: S, Sridhar, et al.
Pubblicazione: (2025)
di: S, Sridhar, et al.
Pubblicazione: (2025)
HiSC4D: Human-centered interaction and 4D Scene Capture in Large-scale Space Using Wearable IMUs and LiDAR
di: Dai, Yudi, et al.
Pubblicazione: (2024)
di: Dai, Yudi, et al.
Pubblicazione: (2024)
Instruction-Driven 3D Facial Expression Generation and Transition
di: Vo, Anh H., et al.
Pubblicazione: (2026)
di: Vo, Anh H., et al.
Pubblicazione: (2026)
A Survey on 3D Gaussian Splatting
di: Chen, Guikun, et al.
Pubblicazione: (2024)
di: Chen, Guikun, et al.
Pubblicazione: (2024)
FlashSplat: 2D to 3D Gaussian Splatting Segmentation Solved Optimally
di: Shen, Qiuhong, et al.
Pubblicazione: (2024)
di: Shen, Qiuhong, et al.
Pubblicazione: (2024)
ANVIL: Analogies and Videos for Lecturers
di: Noviello, Yuri, et al.
Pubblicazione: (2026)
di: Noviello, Yuri, et al.
Pubblicazione: (2026)
On Copyright Risks of Text-to-Image Diffusion Models
di: Zhang, Yang, et al.
Pubblicazione: (2023)
di: Zhang, Yang, et al.
Pubblicazione: (2023)
Photoreal Scene Reconstruction from an Egocentric Device
di: Lv, Zhaoyang, et al.
Pubblicazione: (2025)
di: Lv, Zhaoyang, et al.
Pubblicazione: (2025)
Minecraft-ify: Minecraft Style Image Generation with Text-guided Image Editing for In-Game Application
di: Kim, Bumsoo, et al.
Pubblicazione: (2024)
di: Kim, Bumsoo, et al.
Pubblicazione: (2024)
Freehand Sketch Generation from Mechanical Components
di: Liao, Zhichao, et al.
Pubblicazione: (2024)
di: Liao, Zhichao, et al.
Pubblicazione: (2024)
LAV: Audio-Driven Dynamic Visual Generation with Neural Compression and StyleGAN2
di: Jung, Jongmin, et al.
Pubblicazione: (2025)
di: Jung, Jongmin, et al.
Pubblicazione: (2025)
Seeing World Dynamics in a Nutshell
di: Shen, Qiuhong, et al.
Pubblicazione: (2025)
di: Shen, Qiuhong, et al.
Pubblicazione: (2025)
Instant3D: Instant Text-to-3D Generation
di: Li, Ming, et al.
Pubblicazione: (2023)
di: Li, Ming, et al.
Pubblicazione: (2023)
Neuro-Oracle: A Trajectory-Aware Agentic RAG Framework for Interpretable Epilepsy Surgical Prognosis
di: Aiersilan, Aizierjiang, et al.
Pubblicazione: (2026)
di: Aiersilan, Aizierjiang, et al.
Pubblicazione: (2026)
HiScene: Creating Hierarchical 3D Scenes with Isometric View Generation
di: Dong, Wenqi, et al.
Pubblicazione: (2025)
di: Dong, Wenqi, et al.
Pubblicazione: (2025)
A Customizable Generator for Comic-Style Visual Narrative
di: Chen, Yi-Chun, et al.
Pubblicazione: (2023)
di: Chen, Yi-Chun, et al.
Pubblicazione: (2023)
PointTalk: Audio-Driven Dynamic Lip Point Cloud for 3D Gaussian-based Talking Head Synthesis
di: Xie, Yifan, et al.
Pubblicazione: (2024)
di: Xie, Yifan, et al.
Pubblicazione: (2024)
Towards Unified Co-Speech Gesture Generation via Hierarchical Implicit Periodicity Learning
di: Guo, Xin, et al.
Pubblicazione: (2025)
di: Guo, Xin, et al.
Pubblicazione: (2025)
ChoreoMuse: Robust Music-to-Dance Video Generation with Style Transfer and Beat-Adherent Motion
di: Wang, Xuanchen, et al.
Pubblicazione: (2025)
di: Wang, Xuanchen, et al.
Pubblicazione: (2025)
Zero-Shot Visual Deepfake Detection: Can AI Predict and Prevent Fake Content Before It's Created?
di: Sar, Ayan, et al.
Pubblicazione: (2025)
di: Sar, Ayan, et al.
Pubblicazione: (2025)
KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation
di: Lyu, Tianle, et al.
Pubblicazione: (2025)
di: Lyu, Tianle, et al.
Pubblicazione: (2025)
Lester: rotoscope animation through video object segmentation and tracking
di: Tous, Ruben
Pubblicazione: (2024)
di: Tous, Ruben
Pubblicazione: (2024)
Extreme Compression of Adaptive Neural Images
di: Hoshikawa, Leo, et al.
Pubblicazione: (2024)
di: Hoshikawa, Leo, et al.
Pubblicazione: (2024)
L3GS: Layered 3D Gaussian Splats for Efficient 3D Scene Delivery
di: Tsai, Yi-Zhen, et al.
Pubblicazione: (2025)
di: Tsai, Yi-Zhen, et al.
Pubblicazione: (2025)
DASC: Depth-of-Field Aware Scene Complexity Metric for 3D Visualization on Light Field Display
di: Akbar, Kamran, et al.
Pubblicazione: (2025)
di: Akbar, Kamran, et al.
Pubblicazione: (2025)
Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation
di: He, Lanshan, et al.
Pubblicazione: (2026)
di: He, Lanshan, et al.
Pubblicazione: (2026)
Large Language Models for Computer-Aided Design: A Survey
di: Zhang, Licheng, et al.
Pubblicazione: (2025)
di: Zhang, Licheng, et al.
Pubblicazione: (2025)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
di: Park, Inkyu, et al.
Pubblicazione: (2023)
di: Park, Inkyu, et al.
Pubblicazione: (2023)
Identity Preserving 3D Head Stylization with Multiview Score Distillation
di: Bilecen, Bahri Batuhan, et al.
Pubblicazione: (2024)
di: Bilecen, Bahri Batuhan, et al.
Pubblicazione: (2024)
SMPLer: Taming Transformers for Monocular 3D Human Shape and Pose Estimation
di: Xu, Xiangyu, et al.
Pubblicazione: (2024)
di: Xu, Xiangyu, et al.
Pubblicazione: (2024)
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
di: Sun, Zeyi, et al.
Pubblicazione: (2024)
di: Sun, Zeyi, et al.
Pubblicazione: (2024)
Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
di: Girdhar, Rohit, et al.
Pubblicazione: (2023)
di: Girdhar, Rohit, et al.
Pubblicazione: (2023)
GTLR-GS: Geometry-Texture Aware LiDAR-Regularized 3D Gaussian Splatting for Realistic Scene Reconstruction
di: Fang, Yan, et al.
Pubblicazione: (2026)
di: Fang, Yan, et al.
Pubblicazione: (2026)
DesignAsCode: Bridging Structural Editability and Visual Fidelity in Graphic Design Generation
di: Liu, Ziyuan, et al.
Pubblicazione: (2026)
di: Liu, Ziyuan, et al.
Pubblicazione: (2026)
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
di: Zheng, Shuhong, et al.
Pubblicazione: (2026)
di: Zheng, Shuhong, et al.
Pubblicazione: (2026)
Time-to-Move: Training-Free Motion Controlled Video Generation via Dual-Clock Denoising
di: Singer, Assaf, et al.
Pubblicazione: (2025)
di: Singer, Assaf, et al.
Pubblicazione: (2025)
Coral Model Generation from Single Images for Virtual Reality Applications
di: Fu, Jie, et al.
Pubblicazione: (2024)
di: Fu, Jie, et al.
Pubblicazione: (2024)
Documenti analoghi
-
GameTileNet: A Semantic Dataset for Low-Resolution Game Art in Procedural Content Generation
di: Chen, Yi-Chun, et al.
Pubblicazione: (2025) -
Think2Sing: Orchestrating Structured Motion Subtitles for Singing-Driven 3D Head Animation
di: Huang, Zikai, et al.
Pubblicazione: (2025) -
PediatricsMQA: a Multi-modal Pediatrics Question Answering Benchmark
di: Bahaj, Adil, et al.
Pubblicazione: (2025) -
Multimodal Cinematic Video Synthesis Using Text-to-Image and Audio Generation Models
di: S, Sridhar, et al.
Pubblicazione: (2025) -
HiSC4D: Human-centered interaction and 4D Scene Capture in Large-scale Space Using Wearable IMUs and LiDAR
di: Dai, Yudi, et al.
Pubblicazione: (2024)