Novel Object Synthesis via Adaptive Text-Image Harmony
Fuente:
arXiv
Saved in:
| Main Authors: | Xiong, Zeren, Zhang, Zedong, Chen, Zikun, Chen, Shuo, Li, Xiang, Sun, Gan, Yang, Jian, Li, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VMDiff: Visual Mixing Diffusion for Limitless Cross-Object Synthesis
by: Xiong, Zeren, et al.
Published: (2025)
by: Xiong, Zeren, et al.
Published: (2025)
Category-Aware 3D Object Composition with Disentangled Texture and Shape Multi-view Diffusion
by: Xiong, Zeren, et al.
Published: (2025)
by: Xiong, Zeren, et al.
Published: (2025)
RMLer: Synthesizing Novel Objects across Diverse Categories via Reinforcement Mixing Learning
by: Li, Jun, et al.
Published: (2025)
by: Li, Jun, et al.
Published: (2025)
TP2O: Creative Text Pair-to-Object Generation using Balance Swap-Sampling
by: Li, Jun, et al.
Published: (2023)
by: Li, Jun, et al.
Published: (2023)
AGSwap: Overcoming Category Boundaries in Object Fusion via Adaptive Group Swapping
by: Zhang, Zedong, et al.
Published: (2025)
by: Zhang, Zedong, et al.
Published: (2025)
Self-Creative Text-to-Object Generation using Semantic-Aware Spatial Weighting
by: Yu, Yue, et al.
Published: (2026)
by: Yu, Yue, et al.
Published: (2026)
HQOD: Harmonious Quantization for Object Detection
by: Huang, Long, et al.
Published: (2024)
by: Huang, Long, et al.
Published: (2024)
PlacidDreamer: Advancing Harmony in Text-to-3D Generation
by: Huang, Shuo, et al.
Published: (2024)
by: Huang, Shuo, et al.
Published: (2024)
Harmonious Group Choreography with Trajectory-Controllable Diffusion
by: Dai, Yuqin, et al.
Published: (2024)
by: Dai, Yuqin, et al.
Published: (2024)
YOLO-Count: Differentiable Object Counting for Text-to-Image Generation
by: Zeng, Guanning, et al.
Published: (2025)
by: Zeng, Guanning, et al.
Published: (2025)
Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement
by: Chen, Zhennan, et al.
Published: (2024)
by: Chen, Zhennan, et al.
Published: (2024)
RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution
by: Jian, Siyong, et al.
Published: (2026)
by: Jian, Siyong, et al.
Published: (2026)
Medical Image Synthesis via Fine-Grained Image-Text Alignment and Anatomy-Pathology Prompting
by: Chen, Wenting, et al.
Published: (2024)
by: Chen, Wenting, et al.
Published: (2024)
Harnessing Vision-Language Pretrained Models with Temporal-Aware Adaptation for Referring Video Object Segmentation
by: Zhou, Zikun, et al.
Published: (2024)
by: Zhou, Zikun, et al.
Published: (2024)
Adaptive Image Zoom-in with Bounding Box Transformation for UAV Object Detection
by: Wang, Tao, et al.
Published: (2026)
by: Wang, Tao, et al.
Published: (2026)
Boundary-Guided Camouflaged Object Detection
by: Sun, Yujia, et al.
Published: (2022)
by: Sun, Yujia, et al.
Published: (2022)
Latent Harmony: Synergistic Unified UHD Image Restoration via Latent Space Regularization and Controllable Refinement
by: Liu, Yidi, et al.
Published: (2025)
by: Liu, Yidi, et al.
Published: (2025)
Autoregressive Image Generation Needs Only a Few Lines of Cached Tokens
by: Qin, Ziran, et al.
Published: (2025)
by: Qin, Ziran, et al.
Published: (2025)
AITTI: Learning Adaptive Inclusive Token for Text-to-Image Generation
by: Hou, Xinyu, et al.
Published: (2024)
by: Hou, Xinyu, et al.
Published: (2024)
ERDDCI: Exact Reversible Diffusion via Dual-Chain Inversion for High-Quality Image Editing
by: Dai, Jimin, et al.
Published: (2024)
by: Dai, Jimin, et al.
Published: (2024)
LTOS: Layout-controllable Text-Object Synthesis via Adaptive Cross-attention Fusions
by: Zhao, Xiaoran, et al.
Published: (2024)
by: Zhao, Xiaoran, et al.
Published: (2024)
DreamLight: Towards Harmonious and Consistent Image Relighting
by: Liu, Yong, et al.
Published: (2025)
by: Liu, Yong, et al.
Published: (2025)
Adaptive Guidance Learning for Camouflaged Object Detection
by: Chen, Zhennan, et al.
Published: (2024)
by: Chen, Zhennan, et al.
Published: (2024)
Adaptive FSS: A Novel Few-Shot Segmentation Framework via Prototype Enhancement
by: Wang, Jing, et al.
Published: (2023)
by: Wang, Jing, et al.
Published: (2023)
RoamScene3D: Immersive Text-to-3D Scene Generation via Adaptive Object-aware Roaming
by: Chu, Jisheng, et al.
Published: (2026)
by: Chu, Jisheng, et al.
Published: (2026)
GenHOI: Generalizing Text-driven 4D Human-Object Interaction Synthesis for Unseen Objects
by: Li, Shujia, et al.
Published: (2025)
by: Li, Shujia, et al.
Published: (2025)
Hybrid Classification-Regression Adaptive Loss for Dense Object Detection
by: Huang, Yanquan, et al.
Published: (2024)
by: Huang, Yanquan, et al.
Published: (2024)
PKINet-v2: Towards Powerful and Efficient Poly-Kernel Remote Sensing Object Detection
by: Cai, Xinhao, et al.
Published: (2026)
by: Cai, Xinhao, et al.
Published: (2026)
Modality-Decoupled RGB-Thermal Object Detector via Query Fusion
by: Tian, Chao, et al.
Published: (2026)
by: Tian, Chao, et al.
Published: (2026)
CHITNet: A Complementary to Harmonious Information Transfer Network for Infrared and Visible Image Fusion
by: Du, Keying, et al.
Published: (2023)
by: Du, Keying, et al.
Published: (2023)
High-Fidelity Novel View Synthesis via Splatting-Guided Diffusion
by: Zhang, Xiang, et al.
Published: (2025)
by: Zhang, Xiang, et al.
Published: (2025)
MIGC: Multi-Instance Generation Controller for Text-to-Image Synthesis
by: Zhou, Dewei, et al.
Published: (2024)
by: Zhou, Dewei, et al.
Published: (2024)
Diff-Aid: Inference-time Adaptive Interaction Denoising for Rectified Text-to-Image Generation
by: Li, Binglei, et al.
Published: (2026)
by: Li, Binglei, et al.
Published: (2026)
MotionShot: Adaptive Motion Transfer across Arbitrary Objects for Text-to-Video Generation
by: Liu, Yanchen, et al.
Published: (2025)
by: Liu, Yanchen, et al.
Published: (2025)
SCOUT: Semi-supervised Camouflaged Object Detection by Utilizing Text and Adaptive Data Selection
by: Yan, Weiqi, et al.
Published: (2025)
by: Yan, Weiqi, et al.
Published: (2025)
TIV-Diffusion: Towards Object-Centric Movement for Text-driven Image to Video Generation
by: Wang, Xingrui, et al.
Published: (2024)
by: Wang, Xingrui, et al.
Published: (2024)
Customizing Text-to-Image Diffusion with Object Viewpoint Control
by: Kumari, Nupur, et al.
Published: (2024)
by: Kumari, Nupur, et al.
Published: (2024)
S$^2$Edit: Text-Guided Image Editing with Precise Semantic and Spatial Control
by: Liu, Xudong, et al.
Published: (2025)
by: Liu, Xudong, et al.
Published: (2025)
SAMKD: Spatial-aware Adaptive Masking Knowledge Distillation for Object Detection
by: Zhang, Zhourui, et al.
Published: (2025)
by: Zhang, Zhourui, et al.
Published: (2025)
Noise Diffusion for Enhancing Semantic Faithfulness in Text-to-Image Synthesis
by: Miao, Boming, et al.
Published: (2024)
by: Miao, Boming, et al.
Published: (2024)
Similar Items
-
VMDiff: Visual Mixing Diffusion for Limitless Cross-Object Synthesis
by: Xiong, Zeren, et al.
Published: (2025) -
Category-Aware 3D Object Composition with Disentangled Texture and Shape Multi-view Diffusion
by: Xiong, Zeren, et al.
Published: (2025) -
RMLer: Synthesizing Novel Objects across Diverse Categories via Reinforcement Mixing Learning
by: Li, Jun, et al.
Published: (2025) -
TP2O: Creative Text Pair-to-Object Generation using Balance Swap-Sampling
by: Li, Jun, et al.
Published: (2023) -
AGSwap: Overcoming Category Boundaries in Object Fusion via Adaptive Group Swapping
by: Zhang, Zedong, et al.
Published: (2025)