Consistent text-to-image generation via scene de-contextualization
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Song, Gong, Peihao, Li, Kunyu, Guo, Kai, Wang, Boyu, Ye, Mao, Zhang, Jianwei, Zhu, Xiatian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unified Source-Free Domain Adaptation
by: Tang, Song, et al.
Published: (2024)
by: Tang, Song, et al.
Published: (2024)
Proxy Denoising for Source-Free Domain Adaptation
by: Tang, Song, et al.
Published: (2024)
by: Tang, Song, et al.
Published: (2024)
Source-Free Domain Adaptation with Vision-Language Prior
by: Tang, Song, et al.
Published: (2026)
by: Tang, Song, et al.
Published: (2026)
Source-Free Domain Adaptive Object Detection with Semantics Compensation
by: Tang, Song, et al.
Published: (2024)
by: Tang, Song, et al.
Published: (2024)
Source-Free Domain Adaptation with Frozen Multimodal Foundation Model
by: Tang, Song, et al.
Published: (2023)
by: Tang, Song, et al.
Published: (2023)
Few-Shot Medical Image Segmentation with High-Fidelity Prototypes
by: Tang, Song, et al.
Published: (2024)
by: Tang, Song, et al.
Published: (2024)
Efficient scene text image super-resolution with semantic guidance
by: TomyEnrique, LeoWu, et al.
Published: (2024)
by: TomyEnrique, LeoWu, et al.
Published: (2024)
Trinity Detector:text-assisted and attention mechanisms based spectral fusion for diffusion generation image detection
by: Song, Jiawei, et al.
Published: (2024)
by: Song, Jiawei, et al.
Published: (2024)
FACTOR: Counterfactual Training-Free Test-Time Adaptation for Open-Vocabulary Object Detection
by: Zhao, Kaixiang, et al.
Published: (2026)
by: Zhao, Kaixiang, et al.
Published: (2026)
Is Foreground Prototype Sufficient? Few-Shot Medical Image Segmentation with Background-Fused Prototype
by: Tang, Song, et al.
Published: (2024)
by: Tang, Song, et al.
Published: (2024)
Domain Adaptive Diabetic Retinopathy Grading with Model Absence and Flowing Data
by: Su, Wenxin, et al.
Published: (2024)
by: Su, Wenxin, et al.
Published: (2024)
3D scene generation from scene graphs and self-attention
by: Bonazzi, Pietro, et al.
Published: (2024)
by: Bonazzi, Pietro, et al.
Published: (2024)
DeMo++: Motion Decoupling for Autonomous Driving
by: Zhang, Bozhou, et al.
Published: (2025)
by: Zhang, Bozhou, et al.
Published: (2025)
Motion Forecasting in Continuous Driving
by: Song, Nan, et al.
Published: (2024)
by: Song, Nan, et al.
Published: (2024)
SituationalLLM: Proactive language models with scene awareness for dynamic, contextual task guidance
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
Exploring text-to-image generation for historical document image retrieval
by: Cote, Melissa, et al.
Published: (2025)
by: Cote, Melissa, et al.
Published: (2025)
PointT2I: LLM-based text-to-image generation via keypoints
by: Lee, Taekyung, et al.
Published: (2025)
by: Lee, Taekyung, et al.
Published: (2025)
3MOS: Multi-sources, Multi-resolutions, and Multi-scenes dataset for Optical-SAR image matching
by: Ye, Yibin, et al.
Published: (2024)
by: Ye, Yibin, et al.
Published: (2024)
LMAD: Integrated End-to-End Vision-Language Model for Explainable Autonomous Driving
by: Song, Nan, et al.
Published: (2025)
by: Song, Nan, et al.
Published: (2025)
ConceptHash: Interpretable Fine-Grained Hashing via Concept Discovery
by: Ng, Kam Woh, et al.
Published: (2024)
by: Ng, Kam Woh, et al.
Published: (2024)
AgentPose: Progressive Distribution Alignment via Feature Agent for Human Pose Distillation
by: Zhang, Feng, et al.
Published: (2025)
by: Zhang, Feng, et al.
Published: (2025)
High Dynamic Range 3D Gaussian Splatting via Luminance-Chromaticity Decomposition
by: Zhang, Kaixuan, et al.
Published: (2025)
by: Zhang, Kaixuan, et al.
Published: (2025)
High Dynamic Range Novel View Synthesis with Single Exposure
by: Zhang, Kaixuan, et al.
Published: (2025)
by: Zhang, Kaixuan, et al.
Published: (2025)
RealEngine: Simulating Autonomous Driving in Realistic Context
by: Jiang, Junzhe, et al.
Published: (2025)
by: Jiang, Junzhe, et al.
Published: (2025)
Future-Aware End-to-End Driving: Bidirectional Modeling of Trajectory Planning and Scene Evolution
by: Zhang, Bozhou, et al.
Published: (2025)
by: Zhang, Bozhou, et al.
Published: (2025)
IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation
by: Tiwari, Amritanshu, et al.
Published: (2025)
by: Tiwari, Amritanshu, et al.
Published: (2025)
Visual question answering based evaluation metrics for text-to-image generation
by: Miyamoto, Mizuki, et al.
Published: (2024)
by: Miyamoto, Mizuki, et al.
Published: (2024)
Enhancing the quality of gauge images captured in smoke and haze scenes through deep learning
by: Ramírez-Agudelo, Oscar H., et al.
Published: (2026)
by: Ramírez-Agudelo, Oscar H., et al.
Published: (2026)
Enhancing High-Resolution 3D Generation through Pixel-wise Gradient Clipping
by: Pan, Zijie, et al.
Published: (2023)
by: Pan, Zijie, et al.
Published: (2023)
Robust Low-Light Human Pose Estimation through Illumination-Texture Modulation
by: Zhang, Feng, et al.
Published: (2025)
by: Zhang, Feng, et al.
Published: (2025)
Preconditioned Score-based Generative Models
by: Ma, Hengyuan, et al.
Published: (2023)
by: Ma, Hengyuan, et al.
Published: (2023)
TensoFlow: Tensorial Flow-based Sampler for Inverse Rendering
by: Gu, Chun, et al.
Published: (2025)
by: Gu, Chun, et al.
Published: (2025)
Efficient4D: Fast Dynamic 3D Object Generation from a Single-view Video
by: Pan, Zijie, et al.
Published: (2024)
by: Pan, Zijie, et al.
Published: (2024)
Teaching in adverse scenes: a statistically feedback-driven threshold and mask adjustment teacher-student framework for object detection in UAV images under adverse scenes
by: Chen, Hongyu, et al.
Published: (2025)
by: Chen, Hongyu, et al.
Published: (2025)
See Tomorrow, Act Today: Foresight-Driven Autonomous Driving
by: Zhang, Bozhou, et al.
Published: (2026)
by: Zhang, Bozhou, et al.
Published: (2026)
Bayesian Test-Time Adaptation for Vision-Language Models
by: Zhou, Lihua, et al.
Published: (2025)
by: Zhou, Lihua, et al.
Published: (2025)
Dark Miner: Defend against undesirable generation for text-to-image diffusion models
by: Meng, Zheling, et al.
Published: (2024)
by: Meng, Zheling, et al.
Published: (2024)
Robust Low-light Scene Restoration via Illumination Transition
by: Li, Ze, et al.
Published: (2025)
by: Li, Ze, et al.
Published: (2025)
Indoor scene recognition from images under visual corruptions
by: Costa, Willams de Lima, et al.
Published: (2024)
by: Costa, Willams de Lima, et al.
Published: (2024)
EC-Depth: Exploring the consistency of self-supervised monocular depth estimation in challenging scenes
by: Song, Ziyang, et al.
Published: (2023)
by: Song, Ziyang, et al.
Published: (2023)
Similar Items
-
Unified Source-Free Domain Adaptation
by: Tang, Song, et al.
Published: (2024) -
Proxy Denoising for Source-Free Domain Adaptation
by: Tang, Song, et al.
Published: (2024) -
Source-Free Domain Adaptation with Vision-Language Prior
by: Tang, Song, et al.
Published: (2026) -
Source-Free Domain Adaptive Object Detection with Semantics Compensation
by: Tang, Song, et al.
Published: (2024) -
Source-Free Domain Adaptation with Frozen Multimodal Foundation Model
by: Tang, Song, et al.
Published: (2023)