Event-Driven Storytelling with Multiple Lifelike Humans in a 3D Scene
Fuente:
arXiv
Salvato in:
| Autori principali: | Lim, Donggeun, Bae, Jinseok, Hwang, Inwoo, Lee, Seungmin, Lee, Hwanhee, Kim, Young Min |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Motion Synthesis with Sparse and Flexible Keyjoint Control
di: Hwang, Inwoo, et al.
Pubblicazione: (2025)
di: Hwang, Inwoo, et al.
Pubblicazione: (2025)
EgoForce: Robust Online Egocentric Motion Reconstruction via Diffusion Forcing
di: Hwang, Inwoo, et al.
Pubblicazione: (2026)
di: Hwang, Inwoo, et al.
Pubblicazione: (2026)
SceneMI: Motion In-betweening for Modeling Human-Scene Interactions
di: Hwang, Inwoo, et al.
Pubblicazione: (2025)
di: Hwang, Inwoo, et al.
Pubblicazione: (2025)
Less is More: Improving Motion Diffusion Models with Sparse Keyframes
di: Bae, Jinseok, et al.
Pubblicazione: (2025)
di: Bae, Jinseok, et al.
Pubblicazione: (2025)
LTGS: Long-Term Gaussian Scene Chronology From Sparse View Updates
di: Kim, Minkwan, et al.
Pubblicazione: (2025)
di: Kim, Minkwan, et al.
Pubblicazione: (2025)
Learning 3D Scene Analogies with Neural Contextual Scene Maps
di: Kim, Junho, et al.
Pubblicazione: (2025)
di: Kim, Junho, et al.
Pubblicazione: (2025)
Towards Spatially Consistent Image Generation: On Incorporating Intrinsic Scene Properties into Diffusion Models
di: Lee, Hyundo, et al.
Pubblicazione: (2025)
di: Lee, Hyundo, et al.
Pubblicazione: (2025)
ScaleMoGen: Autoregressive Next-Scale Prediction for Human Motion Generation
di: Hwang, Inwoo, et al.
Pubblicazione: (2026)
di: Hwang, Inwoo, et al.
Pubblicazione: (2026)
Difficulty-Aware Label-Guided Denoising for Monocular 3D Object Detection
di: Lee, Soyul, et al.
Pubblicazione: (2025)
di: Lee, Soyul, et al.
Pubblicazione: (2025)
Finding 3D Scene Analogies with Multimodal Foundation Models
di: Kim, Junho, et al.
Pubblicazione: (2025)
di: Kim, Junho, et al.
Pubblicazione: (2025)
Locality-aware Concept Bottleneck Model
di: Jeon, Sujin, et al.
Pubblicazione: (2025)
di: Jeon, Sujin, et al.
Pubblicazione: (2025)
SnapMoGen: Human Motion Generation from Expressive Texts
di: Guo, Chuan, et al.
Pubblicazione: (2025)
di: Guo, Chuan, et al.
Pubblicazione: (2025)
RoEL: Robust Event-based 3D Line Reconstruction
di: Bae, Gwangtak, et al.
Pubblicazione: (2026)
di: Bae, Gwangtak, et al.
Pubblicazione: (2026)
Versatile Physics-based Character Control with Hybrid Latent Representation
di: Bae, Jinseok, et al.
Pubblicazione: (2025)
di: Bae, Jinseok, et al.
Pubblicazione: (2025)
Event-Driven Dynamic Scene Depth Completion
di: Yan, Zhiqiang, et al.
Pubblicazione: (2025)
di: Yan, Zhiqiang, et al.
Pubblicazione: (2025)
Retaining and Enhancing Pre-trained Knowledge in Vision-Language Models with Prompt Ensembling
di: Kim, Donggeun, et al.
Pubblicazione: (2024)
di: Kim, Donggeun, et al.
Pubblicazione: (2024)
Humans as a Calibration Pattern: Dynamic 3D Scene Reconstruction from Unsynchronized and Uncalibrated Videos
di: Choi, Changwoon, et al.
Pubblicazione: (2024)
di: Choi, Changwoon, et al.
Pubblicazione: (2024)
A Mixture of Experts Approach to 3D Human Motion Prediction
di: Shieh, Edmund, et al.
Pubblicazione: (2024)
di: Shieh, Edmund, et al.
Pubblicazione: (2024)
Towards Temporal Fusion Beyond the Field of View for Camera-based Semantic Scene Completion
di: Bae, Jongseong, et al.
Pubblicazione: (2025)
di: Bae, Jongseong, et al.
Pubblicazione: (2025)
VASA-1: Lifelike Audio-Driven Talking Faces Generated in Real Time
di: Xu, Sicheng, et al.
Pubblicazione: (2024)
di: Xu, Sicheng, et al.
Pubblicazione: (2024)
Decoupled Generative Modeling for Human-Object Interaction Synthesis
di: Jung, Hwanhee, et al.
Pubblicazione: (2025)
di: Jung, Hwanhee, et al.
Pubblicazione: (2025)
VASA-3D: Lifelike Audio-Driven Gaussian Head Avatars from a Single Image
di: Xu, Sicheng, et al.
Pubblicazione: (2025)
di: Xu, Sicheng, et al.
Pubblicazione: (2025)
3Doodle: Compact Abstraction of Objects with 3D Strokes
di: Choi, Changwoon, et al.
Pubblicazione: (2024)
di: Choi, Changwoon, et al.
Pubblicazione: (2024)
MASH-VLM: Mitigating Action-Scene Hallucination in Video-LLMs through Disentangled Spatial-Temporal Representations
di: Bae, Kyungho, et al.
Pubblicazione: (2025)
di: Bae, Kyungho, et al.
Pubblicazione: (2025)
The Era of Foundation Models in Medical Imaging is Approaching : A Scoping Review of the Clinical Value of Large-Scale Generative AI Applications in Radiology
di: Seo, Inwoo, et al.
Pubblicazione: (2024)
di: Seo, Inwoo, et al.
Pubblicazione: (2024)
EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions
di: Chen, Zhiyuan, et al.
Pubblicazione: (2024)
di: Chen, Zhiyuan, et al.
Pubblicazione: (2024)
CanonicalFusion: Generating Drivable 3D Human Avatars from Multiple Images
di: Shin, Jisu, et al.
Pubblicazione: (2024)
di: Shin, Jisu, et al.
Pubblicazione: (2024)
Recovering Dynamic 3D Sketches from Videos
di: Lee, Jaeah, et al.
Pubblicazione: (2025)
di: Lee, Jaeah, et al.
Pubblicazione: (2025)
3D-LLDM: Label-Guided 3D Latent Diffusion Model for Improving High-Resolution Synthetic MR Imaging in Hepatic Structure Segmentation
di: Kim, Kyeonghun, et al.
Pubblicazione: (2026)
di: Kim, Kyeonghun, et al.
Pubblicazione: (2026)
PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
Integrating Meshes and 3D Gaussians for Indoor Scene Reconstruction with SAM Mask Guidance
di: Kim, Jiyeop, et al.
Pubblicazione: (2024)
di: Kim, Jiyeop, et al.
Pubblicazione: (2024)
VEGS: View Extrapolation of Urban Scenes in 3D Gaussian Splatting using Learned Priors
di: Hwang, Sungwon, et al.
Pubblicazione: (2024)
di: Hwang, Sungwon, et al.
Pubblicazione: (2024)
Missing Modality Prediction for Unpaired Multimodal Learning via Joint Embedding of Unimodal Models
di: Kim, Donggeun, et al.
Pubblicazione: (2024)
di: Kim, Donggeun, et al.
Pubblicazione: (2024)
ReMP: Reusable Motion Prior for Multi-domain 3D Human Pose Estimation and Motion Inbetweening
di: Jang, Hojun, et al.
Pubblicazione: (2024)
di: Jang, Hojun, et al.
Pubblicazione: (2024)
Guess The Unseen: Dynamic 3D Scene Reconstruction from Partial 2D Glimpses
di: Lee, Inhee, et al.
Pubblicazione: (2024)
di: Lee, Inhee, et al.
Pubblicazione: (2024)
MoRGS: Efficient Per-Gaussian Motion Reasoning for Streamable Dynamic 3D Scenes
di: Lee, Wonjoon, et al.
Pubblicazione: (2026)
di: Lee, Wonjoon, et al.
Pubblicazione: (2026)
MORDA: A Synthetic Dataset to Facilitate Adaptation of Object Detectors to Unseen Real-target Domain While Preserving Performance on Real-source Domain
di: Lim, Hojun, et al.
Pubblicazione: (2025)
di: Lim, Hojun, et al.
Pubblicazione: (2025)
PASTA: Part-Aware Sketch-to-3D Shape Generation with Text-Aligned Prior
di: Lee, Seunggwan, et al.
Pubblicazione: (2025)
di: Lee, Seunggwan, et al.
Pubblicazione: (2025)
Analogical Trajectory Transfer
di: Kim, Junho, et al.
Pubblicazione: (2026)
di: Kim, Junho, et al.
Pubblicazione: (2026)
Pushing the Limits of Vision-Language Models in Remote Sensing without Human Annotations
di: Cha, Keumgang, et al.
Pubblicazione: (2024)
di: Cha, Keumgang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Motion Synthesis with Sparse and Flexible Keyjoint Control
di: Hwang, Inwoo, et al.
Pubblicazione: (2025) -
EgoForce: Robust Online Egocentric Motion Reconstruction via Diffusion Forcing
di: Hwang, Inwoo, et al.
Pubblicazione: (2026) -
SceneMI: Motion In-betweening for Modeling Human-Scene Interactions
di: Hwang, Inwoo, et al.
Pubblicazione: (2025) -
Less is More: Improving Motion Diffusion Models with Sparse Keyframes
di: Bae, Jinseok, et al.
Pubblicazione: (2025) -
LTGS: Long-Term Gaussian Scene Chronology From Sparse View Updates
di: Kim, Minkwan, et al.
Pubblicazione: (2025)