BeyondScene: Higher-Resolution Human-Centric Scene Generation With Pretrained Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Gwanghyun, Kim, Hayeon, Seo, Hoigi, Kang, Dong Un, Chun, Se Young |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion
by: Kim, Gwanghyun, et al.
Published: (2024)
by: Kim, Gwanghyun, et al.
Published: (2024)
LaMoGen: Laban Movement-Guided Diffusion for Text-to-Motion Generation
by: Kim, Heechang, et al.
Published: (2025)
by: Kim, Heechang, et al.
Published: (2025)
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
by: Seo, Hoigi, et al.
Published: (2025)
by: Seo, Hoigi, et al.
Published: (2025)
Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models
by: Jeong, Wongi, et al.
Published: (2026)
by: Jeong, Wongi, et al.
Published: (2026)
Efficient Personalization of Quantized Diffusion Model without Backpropagation
by: Seo, Hoigi, et al.
Published: (2025)
by: Seo, Hoigi, et al.
Published: (2025)
Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation
by: Seo, Hoigi, et al.
Published: (2025)
by: Seo, Hoigi, et al.
Published: (2025)
Training-free Mixed-Resolution Latent Upsampling for Spatially Accelerated Diffusion Transformers
by: Jeong, Wongi, et al.
Published: (2025)
by: Jeong, Wongi, et al.
Published: (2025)
Contribution-based Low-Rank Adaptation with Pre-training Model for Real Image Restoration
by: Park, Donwon, et al.
Published: (2024)
by: Park, Donwon, et al.
Published: (2024)
Human Interaction-Aware 3D Reconstruction from a Single Image
by: Kim, Gwanghyun, et al.
Published: (2026)
by: Kim, Gwanghyun, et al.
Published: (2026)
Erasing Thousands of Concepts: Towards Scalable and Practical Concept Erasure for Text-to-Image Diffusion Models
by: Seo, Hoigi, et al.
Published: (2026)
by: Seo, Hoigi, et al.
Published: (2026)
Robust 3D-Masked Part-level Editing in 3D Gaussian Splatting with Regularized Score Distillation Sampling
by: Kim, Hayeon, et al.
Published: (2025)
by: Kim, Hayeon, et al.
Published: (2025)
Self-Cascaded Diffusion Models for Arbitrary-Scale Image Super-Resolution
by: Bang, Junseo, et al.
Published: (2025)
by: Bang, Junseo, et al.
Published: (2025)
INTRA: Interaction Relationship-aware Weakly Supervised Affordance Grounding
by: Jang, Ji Ha, et al.
Published: (2024)
by: Jang, Ji Ha, et al.
Published: (2024)
Short-term Object Interaction Anticipation with Disentangled Object Detection @ Ego4D Short Term Object Interaction Anticipation Challenge
by: Cho, Hyunjin, et al.
Published: (2024)
by: Cho, Hyunjin, et al.
Published: (2024)
Triadic Dynamics Aware Diffusion Posterior Sampling for Inverse Problems: Optimizing Guidance and Stochasticity Schedules
by: Bang, Junseo, et al.
Published: (2026)
by: Bang, Junseo, et al.
Published: (2026)
Geometrical Properties of Text Token Embeddings for Strong Semantic Binding in Text-to-Image Generation
by: Seo, Hoigi, et al.
Published: (2025)
by: Seo, Hoigi, et al.
Published: (2025)
Uncertainty-guided Compositional Alignment with Part-to-Whole Semantic Representativeness in Hyperbolic Vision-Language Models
by: Kim, Hayeon, et al.
Published: (2026)
by: Kim, Hayeon, et al.
Published: (2026)
SceneAdapt: Scene-aware Adaptation of Human Motion Diffusion
by: Cho, Jungbin, et al.
Published: (2025)
by: Cho, Jungbin, et al.
Published: (2025)
GeoMan: Temporally Consistent Human Geometry Estimation using Image-to-Video Diffusion
by: Kim, Gwanghyun, et al.
Published: (2025)
by: Kim, Gwanghyun, et al.
Published: (2025)
Look Beyond: Two-Stage Scene View Generation via Panorama and Video Diffusion
by: Kang, Xueyang, et al.
Published: (2025)
by: Kang, Xueyang, et al.
Published: (2025)
SceneLinker: Compositional 3D Scene Generation via Semantic Scene Graph from RGB Sequences
by: Kim, Seok-Young, et al.
Published: (2026)
by: Kim, Seok-Young, et al.
Published: (2026)
Towards Generalizable Scene Change Detection
by: Kim, Jaewoo, et al.
Published: (2024)
by: Kim, Jaewoo, et al.
Published: (2024)
DualTSR: Unified Dual-Diffusion Transformer for Scene Text Image Super-Resolution
by: Niu, Axi, et al.
Published: (2026)
by: Niu, Axi, et al.
Published: (2026)
SceneMI: Motion In-betweening for Modeling Human-Scene Interactions
by: Hwang, Inwoo, et al.
Published: (2025)
by: Hwang, Inwoo, et al.
Published: (2025)
HetScene: Heterogeneity-Aware Diffusion for Dense Indoor Scene Generation
by: Chen, Zini, et al.
Published: (2026)
by: Chen, Zini, et al.
Published: (2026)
How Blind and Low-Vision Individuals Prefer Large Vision-Language Model-Generated Scene Descriptions
by: An, Na Min, et al.
Published: (2025)
by: An, Na Min, et al.
Published: (2025)
InsertAnywhere: Bridging 4D Scene Geometry and Diffusion Models for Realistic Video Object Insertion
by: Jin, Hoiyeong, et al.
Published: (2025)
by: Jin, Hoiyeong, et al.
Published: (2025)
Adaptive Self-training Framework for Fine-grained Scene Graph Generation
by: Kim, Kibum, et al.
Published: (2024)
by: Kim, Kibum, et al.
Published: (2024)
Concept Pinpoint Eraser for Text-to-image Diffusion Models via Residual Attention Gate
by: Lee, Byung Hyun, et al.
Published: (2025)
by: Lee, Byung Hyun, et al.
Published: (2025)
EchoScene: Indoor Scene Generation via Information Echo over Scene Graph Diffusion
by: Zhai, Guangyao, et al.
Published: (2024)
by: Zhai, Guangyao, et al.
Published: (2024)
Scene Graph Generation Strategy with Co-occurrence Knowledge and Learnable Term Frequency
by: Kim, Hyeongjin, et al.
Published: (2024)
by: Kim, Hyeongjin, et al.
Published: (2024)
Diffusion for Out-of-Distribution Detection on Road Scenes and Beyond
by: Galesso, Silvio, et al.
Published: (2024)
by: Galesso, Silvio, et al.
Published: (2024)
DOCTR: Disentangled Object-Centric Transformer for Point Scene Understanding
by: Yu, Xiaoxuan, et al.
Published: (2024)
by: Yu, Xiaoxuan, et al.
Published: (2024)
Scene Splatter: Momentum 3D Scene Generation from Single Image with Video Diffusion Model
by: Zhang, Shengjun, et al.
Published: (2025)
by: Zhang, Shengjun, et al.
Published: (2025)
SSEditor: Controllable Mask-to-Scene Generation with Diffusion Model
by: Zheng, Haowen, et al.
Published: (2024)
by: Zheng, Haowen, et al.
Published: (2024)
EditCrafter: Tuning-free High-Resolution Image Editing via Pretrained Diffusion Model
by: Kim, Kunho, et al.
Published: (2026)
by: Kim, Kunho, et al.
Published: (2026)
Recasting Generic Pretrained Vision Transformers As Object-Centric Scene Encoders For Manipulation Policies
by: Qian, Jianing, et al.
Published: (2024)
by: Qian, Jianing, et al.
Published: (2024)
Latent Space Super-Resolution for Higher-Resolution Image Generation with Diffusion Models
by: Jeong, Jinho, et al.
Published: (2025)
by: Jeong, Jinho, et al.
Published: (2025)
Finding 3D Scene Analogies with Multimodal Foundation Models
by: Kim, Junho, et al.
Published: (2025)
by: Kim, Junho, et al.
Published: (2025)
Localized Concept Erasure for Text-to-Image Diffusion Models Using Training-Free Gated Low-Rank Adaptation
by: Lee, Byung Hyun, et al.
Published: (2025)
by: Lee, Byung Hyun, et al.
Published: (2025)
Similar Items
-
PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion
by: Kim, Gwanghyun, et al.
Published: (2024) -
LaMoGen: Laban Movement-Guided Diffusion for Text-to-Motion Generation
by: Kim, Heechang, et al.
Published: (2025) -
On Epistemic Uncertainty of Visual Tokens for Object Hallucinations in Large Vision-Language Models
by: Seo, Hoigi, et al.
Published: (2025) -
Training-free, Perceptually Consistent Low-Resolution Previews with High-Resolution Image for Efficient Workflows of Diffusion Models
by: Jeong, Wongi, et al.
Published: (2026) -
Efficient Personalization of Quantized Diffusion Model without Backpropagation
by: Seo, Hoigi, et al.
Published: (2025)