Self-Creative Text-to-Object Generation using Semantic-Aware Spatial Weighting
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Yue, Chen, Haibo, Chen, Shuo, Yang, Jian, Li, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TP2O: Creative Text Pair-to-Object Generation using Balance Swap-Sampling
by: Li, Jun, et al.
Published: (2023)
by: Li, Jun, et al.
Published: (2023)
RMLer: Synthesizing Novel Objects across Diverse Categories via Reinforcement Mixing Learning
by: Li, Jun, et al.
Published: (2025)
by: Li, Jun, et al.
Published: (2025)
VMDiff: Visual Mixing Diffusion for Limitless Cross-Object Synthesis
by: Xiong, Zeren, et al.
Published: (2025)
by: Xiong, Zeren, et al.
Published: (2025)
Novel Object Synthesis via Adaptive Text-Image Harmony
by: Xiong, Zeren, et al.
Published: (2024)
by: Xiong, Zeren, et al.
Published: (2024)
Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement
by: Chen, Zhennan, et al.
Published: (2024)
by: Chen, Zhennan, et al.
Published: (2024)
Vision-Language Models as Differentiable Semantic and Spatial Rewards for Text-to-3D Generation
by: Bai, Weimin, et al.
Published: (2025)
by: Bai, Weimin, et al.
Published: (2025)
Beyond Flat Text: Dual Self-inherited Guidance for Visual Text Generation
by: Luo, Minxing, et al.
Published: (2025)
by: Luo, Minxing, et al.
Published: (2025)
Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking
by: Hu, Xiantao, et al.
Published: (2024)
by: Hu, Xiantao, et al.
Published: (2024)
Semantics-Guided Moving Object Segmentation with 3D LiDAR
by: Gu, Shuo, et al.
Published: (2022)
by: Gu, Shuo, et al.
Published: (2022)
Multi-Object Advertisement Creative Generation
by: Gao, Jialu, et al.
Published: (2026)
by: Gao, Jialu, et al.
Published: (2026)
Improving Text-guided Object Inpainting with Semantic Pre-inpainting
by: Chen, Yifu, et al.
Published: (2024)
by: Chen, Yifu, et al.
Published: (2024)
Text-promptable Object Counting via Quantity Awareness Enhancement
by: Shi, Miaojing, et al.
Published: (2025)
by: Shi, Miaojing, et al.
Published: (2025)
Category-Aware 3D Object Composition with Disentangled Texture and Shape Multi-view Diffusion
by: Xiong, Zeren, et al.
Published: (2025)
by: Xiong, Zeren, et al.
Published: (2025)
Redefining <Creative> in Dictionary: Towards an Enhanced Semantic Understanding of Creative Generation
by: Feng, Fu, et al.
Published: (2024)
by: Feng, Fu, et al.
Published: (2024)
SFFR: Spatial-Frequency Feature Reconstruction for Multispectral Aerial Object Detection
by: Zuo, Xin, et al.
Published: (2025)
by: Zuo, Xin, et al.
Published: (2025)
SAPNet++: Evolving Point-Prompted Instance Segmentation with Semantic and Spatial Awareness
by: Wei, Zhaoyang, et al.
Published: (2026)
by: Wei, Zhaoyang, et al.
Published: (2026)
Exploring Boundary-Aware Spatial-Frequency Fusion for Camouflaged Object Detection
by: Yu, Song, et al.
Published: (2026)
by: Yu, Song, et al.
Published: (2026)
TSdetector: Temporal-Spatial Self-correction Collaborative Learning for Colonoscopy Video Detection
by: Wang, Kaini, et al.
Published: (2024)
by: Wang, Kaini, et al.
Published: (2024)
RTGen: Generating Region-Text Pairs for Open-Vocabulary Object Detection
by: Chen, Fangyi, et al.
Published: (2024)
by: Chen, Fangyi, et al.
Published: (2024)
FSDETR: Frequency-Spatial Feature Enhancement for Small Object Detection
by: Huang, Jianchao, et al.
Published: (2026)
by: Huang, Jianchao, et al.
Published: (2026)
TIGaussian: Disentangle Gaussians for Spatial-Awared Text-Image-3D Alignment
by: Liu, Jiarun, et al.
Published: (2026)
by: Liu, Jiarun, et al.
Published: (2026)
FedSC: Federated Learning with Semantic-Aware Collaboration
by: Wang, Huan, et al.
Published: (2025)
by: Wang, Huan, et al.
Published: (2025)
Think-Then-Generate: Reasoning-Aware Text-to-Image Diffusion with LLM Encoders
by: Kou, Siqi, et al.
Published: (2026)
by: Kou, Siqi, et al.
Published: (2026)
SRSR: Enhancing Semantic Accuracy in Real-World Image Super-Resolution with Spatially Re-Focused Text-Conditioning
by: Chen, Chen, et al.
Published: (2025)
by: Chen, Chen, et al.
Published: (2025)
Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens
by: Kim, Dongwon, et al.
Published: (2025)
by: Kim, Dongwon, et al.
Published: (2025)
SSD: Spatial-Semantic Head Decoupling for Efficient Autoregressive Image Generation
by: Jian, Siyong, et al.
Published: (2025)
by: Jian, Siyong, et al.
Published: (2025)
SMILEtrack: SiMIlarity LEarning for Occlusion-Aware Multiple Object Tracking
by: Wang, Yu-Hsiang, et al.
Published: (2022)
by: Wang, Yu-Hsiang, et al.
Published: (2022)
Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation
by: Ng, Kam Woh, et al.
Published: (2025)
by: Ng, Kam Woh, et al.
Published: (2025)
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
SpatialFormer: Semantic and Target Aware Attentions for Few-Shot Learning
by: Lai, Jinxiang, et al.
Published: (2023)
by: Lai, Jinxiang, et al.
Published: (2023)
DreamMesh: Jointly Manipulating and Texturing Triangle Meshes for Text-to-3D Generation
by: Yang, Haibo, et al.
Published: (2024)
by: Yang, Haibo, et al.
Published: (2024)
TOUCH: Text-guided Controllable Generation of Free-Form Hand-Object Interactions
by: Han, Guangyi, et al.
Published: (2025)
by: Han, Guangyi, et al.
Published: (2025)
Spatially-Weighted CLIP for Street-View Geo-localization
by: Han, Ting, et al.
Published: (2026)
by: Han, Ting, et al.
Published: (2026)
Sparse VideoGen2: Accelerate Video Generation with Sparse Attention via Semantic-Aware Permutation
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
Scale-Aware UAV-to-Satellite Cross-View Geo-Localization: A Semantic Geometric Approach
by: Ye, Yibin, et al.
Published: (2026)
by: Ye, Yibin, et al.
Published: (2026)
FAFA: Frequency-Aware Flow-Aided Self-Supervision for Underwater Object Pose Estimation
by: Tang, Jingyi, et al.
Published: (2024)
by: Tang, Jingyi, et al.
Published: (2024)
Visual Semantic Description Generation with MLLMs for Image-Text Matching
by: Chen, Junyu, et al.
Published: (2025)
by: Chen, Junyu, et al.
Published: (2025)
GenHOI: Generalized Hand-Object Pose Estimation with Occlusion Awareness
by: Yang, Hui, et al.
Published: (2026)
by: Yang, Hui, et al.
Published: (2026)
CritiFusion: Semantic Critique and Spectral Alignment for Faithful Text-to-Image Generation
by: Chen, ZhenQi, et al.
Published: (2025)
by: Chen, ZhenQi, et al.
Published: (2025)
Image-Text Co-Decomposition for Text-Supervised Semantic Segmentation
by: Wu, Ji-Jia, et al.
Published: (2024)
by: Wu, Ji-Jia, et al.
Published: (2024)
Similar Items
-
TP2O: Creative Text Pair-to-Object Generation using Balance Swap-Sampling
by: Li, Jun, et al.
Published: (2023) -
RMLer: Synthesizing Novel Objects across Diverse Categories via Reinforcement Mixing Learning
by: Li, Jun, et al.
Published: (2025) -
VMDiff: Visual Mixing Diffusion for Limitless Cross-Object Synthesis
by: Xiong, Zeren, et al.
Published: (2025) -
Novel Object Synthesis via Adaptive Text-Image Harmony
by: Xiong, Zeren, et al.
Published: (2024) -
Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement
by: Chen, Zhennan, et al.
Published: (2024)