SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Guibao, Wang, Luozhou, Lin, Jiantao, Ge, Wenhang, Zhang, Chaozhe, Tao, Xin, Zhang, Yuan, Wan, Pengfei, Wang, Zhongyuan, Chen, Guangyong, Li, Yijun, Chen, Ying-Cong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Text-Anchored Score Composition: Tackling Condition Misalignment in Text-to-Image Diffusion Models
by: Wang, Luozhou, et al.
Published: (2023)
by: Wang, Luozhou, et al.
Published: (2023)
CamPilot: Improving Camera Control in Video Diffusion Model with Efficient Camera Reward Feedback
by: Ge, Wenhang, et al.
Published: (2026)
by: Ge, Wenhang, et al.
Published: (2026)
PRM: Photometric Stereo based Large Reconstruction Model
by: Ge, Wenhang, et al.
Published: (2024)
by: Ge, Wenhang, et al.
Published: (2024)
Motion Inversion for Video Customization
by: Wang, Luozhou, et al.
Published: (2024)
by: Wang, Luozhou, et al.
Published: (2024)
StereoPilot: Learning Unified and Efficient Stereo Conversion via Generative Priors
by: Shen, Guibao, et al.
Published: (2025)
by: Shen, Guibao, et al.
Published: (2025)
A Mechanistic View on Video Generation as World Models: State and Dynamics
by: Wang, Luozhou, et al.
Published: (2026)
by: Wang, Luozhou, et al.
Published: (2026)
RectifiedHR: Enable Efficient High-Resolution Synthesis via Energy Rectification
by: Yang, Zhen, et al.
Published: (2025)
by: Yang, Zhen, et al.
Published: (2025)
FlexGen: Flexible Multi-View Generation from Text and Image Inputs
by: Xu, Xinli, et al.
Published: (2024)
by: Xu, Xinli, et al.
Published: (2024)
TransPixeler: Advancing Text-to-Video Generation with Transparency
by: Wang, Luozhou, et al.
Published: (2025)
by: Wang, Luozhou, et al.
Published: (2025)
ComfyMind: Toward General-Purpose Generation via Tree-Based Planning and Reactive Feedback
by: Guo, Litao, et al.
Published: (2025)
by: Guo, Litao, et al.
Published: (2025)
SG-Reg: Generalizable and Efficient Scene Graph Registration
by: Liu, Chuhao, et al.
Published: (2025)
by: Liu, Chuhao, et al.
Published: (2025)
Uni-Renderer: Unifying Rendering and Inverse Rendering Via Dual Stream Diffusion
by: Chen, Zhifei, et al.
Published: (2024)
by: Chen, Zhifei, et al.
Published: (2024)
TeSG: Textual Semantic Guidance for Infrared and Visible Image Fusion
by: Zhu, Mingrui, et al.
Published: (2025)
by: Zhu, Mingrui, et al.
Published: (2025)
RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data
by: Chen, Harold Haodong, et al.
Published: (2026)
by: Chen, Harold Haodong, et al.
Published: (2026)
SG-NeRF: Neural Surface Reconstruction with Scene Graph Optimization
by: Chen, Yiyang, et al.
Published: (2024)
by: Chen, Yiyang, et al.
Published: (2024)
LLM-Optic: Unveiling the Capabilities of Large Language Models for Universal Visual Grounding
by: Zhao, Haoyu, et al.
Published: (2024)
by: Zhao, Haoyu, et al.
Published: (2024)
Can We Build Scene Graphs, Not Classify Them? FlowSG: Progressive Image-Conditioned Scene Graph Generation with Flow Matching
by: Hu, Xin, et al.
Published: (2026)
by: Hu, Xin, et al.
Published: (2026)
NeuSG: Neural Implicit Surface Reconstruction with 3D Gaussian Splatting Guidance
by: Chen, Hanlin, et al.
Published: (2023)
by: Chen, Hanlin, et al.
Published: (2023)
DiscoSG: Towards Discourse-Level Text Scene Graph Parsing through Iterative Graph Refinement
by: Lin, Shaoqing, et al.
Published: (2025)
by: Lin, Shaoqing, et al.
Published: (2025)
ArtiSG: Functional 3D Scene Graph Construction via Human-demonstrated Articulated Objects Manipulation
by: Gu, Qiuyi, et al.
Published: (2025)
by: Gu, Qiuyi, et al.
Published: (2025)
Robust SG-NeRF: Robust Scene Graph Aided Neural Surface Reconstruction
by: Gu, Yi, et al.
Published: (2024)
by: Gu, Yi, et al.
Published: (2024)
INHerit-SG: Incremental Hierarchical Semantic Scene Graphs with RAG-Style Retrieval
by: Fang, YukTungSamuel, et al.
Published: (2026)
by: Fang, YukTungSamuel, et al.
Published: (2026)
SG-Tailor: Inter-Object Commonsense Relationship Reasoning for Scene Graph Manipulation
by: Shang, Haoliang, et al.
Published: (2025)
by: Shang, Haoliang, et al.
Published: (2025)
GeoSceneGraph: Geometric Scene Graph Diffusion Model for Text-guided 3D Indoor Scene Synthesis
by: Ruiz, Antonio, et al.
Published: (2025)
by: Ruiz, Antonio, et al.
Published: (2025)
DisEnvisioner: Disentangled and Enriched Visual Prompt for Customized Image Generation
by: He, Jing, et al.
Published: (2024)
by: He, Jing, et al.
Published: (2024)
Restore Text First, Enhance Image Later: Two-Stage Scene Text Image Super-Resolution with Glyph Structure Guidance
by: Luo, Minxing, et al.
Published: (2025)
by: Luo, Minxing, et al.
Published: (2025)
PSGS: Text-driven Panorama Sliding Scene Generation via Gaussian Splatting
by: Zhang, Xin, et al.
Published: (2026)
by: Zhang, Xin, et al.
Published: (2026)
StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style Adapter
by: Liu, Gongye, et al.
Published: (2023)
by: Liu, Gongye, et al.
Published: (2023)
4D Driving Scene Generation With Stereo Forcing
by: Lu, Hao, et al.
Published: (2025)
by: Lu, Hao, et al.
Published: (2025)
Optimal Trajectory‐Following Guidance Based on Receding Horizon Indirect Gauss Pseudospectral Method for a Gliding Aerial Vehicle
by: Qi Chen, et al.
Published: (2026)
by: Qi Chen, et al.
Published: (2026)
STK-Adapter: Incorporating Evolving Graph and Event Chain for Temporal Knowledge Graph Extrapolation
by: Zhao, Shuyuan, et al.
Published: (2026)
by: Zhao, Shuyuan, et al.
Published: (2026)
LLaVA-SG: Leveraging Scene Graphs as Visual Semantic Expression in Vision-Language Models
by: Wang, Jingyi, et al.
Published: (2024)
by: Wang, Jingyi, et al.
Published: (2024)
DiMeR: Disentangled Mesh Reconstruction Model
by: Jiang, Lutao, et al.
Published: (2025)
by: Jiang, Lutao, et al.
Published: (2025)
Probe and Skip: Self-Predictive Token Skipping for Efficient Long-Context LLM Inference
by: Wu, Zimeng, et al.
Published: (2026)
by: Wu, Zimeng, et al.
Published: (2026)
Private Information Retrieval over Graphs
by: Ge, Gennian, et al.
Published: (2025)
by: Ge, Gennian, et al.
Published: (2025)
KeySG: Hierarchical Keyframe-Based 3D Scene Graphs
by: Werby, Abdelrhman, et al.
Published: (2025)
by: Werby, Abdelrhman, et al.
Published: (2025)
CoPa-SG: Dense Scene Graphs with Parametric and Proto-Relations
by: Lorenz, Julian, et al.
Published: (2025)
by: Lorenz, Julian, et al.
Published: (2025)
FlexPainter: Flexible and Multi-View Consistent Texture Generation
by: Yan, Dongyu, et al.
Published: (2025)
by: Yan, Dongyu, et al.
Published: (2025)
How Much To Guide: Revisiting Adaptive Guidance in Classifier-Free Guidance Text-to-Vision Diffusion Models
by: Zhang, Huixuan, et al.
Published: (2025)
by: Zhang, Huixuan, et al.
Published: (2025)
Bi-TTA: Bidirectional Test-Time Adapter for Remote Physiological Measurement
by: Li, Haodong, et al.
Published: (2024)
by: Li, Haodong, et al.
Published: (2024)
Similar Items
-
Text-Anchored Score Composition: Tackling Condition Misalignment in Text-to-Image Diffusion Models
by: Wang, Luozhou, et al.
Published: (2023) -
CamPilot: Improving Camera Control in Video Diffusion Model with Efficient Camera Reward Feedback
by: Ge, Wenhang, et al.
Published: (2026) -
PRM: Photometric Stereo based Large Reconstruction Model
by: Ge, Wenhang, et al.
Published: (2024) -
Motion Inversion for Video Customization
by: Wang, Luozhou, et al.
Published: (2024) -
StereoPilot: Learning Unified and Efficient Stereo Conversion via Generative Priors
by: Shen, Guibao, et al.
Published: (2025)