Compass Control: Multi Object Orientation Control for Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Parihar, Rishubh, Agrawal, Vaibhav, VS, Sachidanand, Babu, R. Venkatesh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PreciseControl: Enhancing Text-To-Image Diffusion Models with Fine-Grained Attribute Control
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
Text2Place: Affordance-aware Text Guided Human Placement
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
SeeThrough3D: Occlusion Aware 3D Control in Text-to-Image Generation
von: Agrawal, Vaibhav, et al.
Veröffentlicht: (2026)
von: Agrawal, Vaibhav, et al.
Veröffentlicht: (2026)
MonoPlace3D: Learning 3D-Aware Object Placement for 3D Monocular Detection
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
Kontinuous Kontext: Continuous Strength Control for Instruction-based Image Editing
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)
Reflecting Reality: Enabling Diffusion Models to Produce Faithful Mirror Reflections
von: Dhiman, Ankit, et al.
Veröffentlicht: (2024)
von: Dhiman, Ankit, et al.
Veröffentlicht: (2024)
Balancing Act: Distribution-Guided Debiasing in Diffusion Models
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024)
Harnessing Diffusion-Generated Synthetic Images for Fair Image Classification
von: Basu, Abhipsa, et al.
Veröffentlicht: (2025)
von: Basu, Abhipsa, et al.
Veröffentlicht: (2025)
Object-Attribute Binding in Text-to-Image Generation: Evaluation and Control
von: Trusca, Maria Mihaela, et al.
Veröffentlicht: (2024)
von: Trusca, Maria Mihaela, et al.
Veröffentlicht: (2024)
GeoDiv: Framework For Measuring Geographical Diversity In Text-To-Image Models
von: Basu, Abhipsa, et al.
Veröffentlicht: (2026)
von: Basu, Abhipsa, et al.
Veröffentlicht: (2026)
Do Vision Language Models Need to Process Image Tokens?
von: Ghosh, Sambit, et al.
Veröffentlicht: (2026)
von: Ghosh, Sambit, et al.
Veröffentlicht: (2026)
Leveraging Vision-Language Models for Improving Domain Generalization in Image Classification
von: Addepalli, Sravanti, et al.
Veröffentlicht: (2023)
von: Addepalli, Sravanti, et al.
Veröffentlicht: (2023)
Customizing Text-to-Image Diffusion with Object Viewpoint Control
von: Kumari, Nupur, et al.
Veröffentlicht: (2024)
von: Kumari, Nupur, et al.
Veröffentlicht: (2024)
OrientDream: Streamlining Text-to-3D Generation with Explicit Orientation Control
von: Huang, Yuzhong, et al.
Veröffentlicht: (2024)
von: Huang, Yuzhong, et al.
Veröffentlicht: (2024)
MIGC: Multi-Instance Generation Controller for Text-to-Image Synthesis
von: Zhou, Dewei, et al.
Veröffentlicht: (2024)
von: Zhou, Dewei, et al.
Veröffentlicht: (2024)
Multitwine: Multi-Object Compositing with Text and Layout Control
von: Tarrés, Gemma Canet, et al.
Veröffentlicht: (2025)
von: Tarrés, Gemma Canet, et al.
Veröffentlicht: (2025)
Robustness Analysis on Foundational Segmentation Models
von: Schiappa, Madeline Chantry, et al.
Veröffentlicht: (2023)
von: Schiappa, Madeline Chantry, et al.
Veröffentlicht: (2023)
OLAF: A Plug-and-Play Framework for Enhanced Multi-object Multi-part Scene Parsing
von: Gupta, Pranav, et al.
Veröffentlicht: (2024)
von: Gupta, Pranav, et al.
Veröffentlicht: (2024)
Objects in Generated Videos Are Slower Than They Appear: Models Suffer Sub-Earth Gravity and Don't Know Galileo's Principle...for now
von: Thozhiyoor, Varun Varma, et al.
Veröffentlicht: (2025)
von: Thozhiyoor, Varun Varma, et al.
Veröffentlicht: (2025)
DOS: Directional Object Separation in Text Embeddings for Multi-Object Image Generation
von: Byun, Dongnam, et al.
Veröffentlicht: (2025)
von: Byun, Dongnam, et al.
Veröffentlicht: (2025)
FOCUS: Optimal Control for Multi-Entity World Modeling in Text-to-Image Generation
von: Bill, Eric Tillmann, et al.
Veröffentlicht: (2025)
von: Bill, Eric Tillmann, et al.
Veröffentlicht: (2025)
MULAN: A Multi Layer Annotated Dataset for Controllable Text-to-Image Generation
von: Tudosiu, Petru-Daniel, et al.
Veröffentlicht: (2024)
von: Tudosiu, Petru-Daniel, et al.
Veröffentlicht: (2024)
Composing Parts for Expressive Object Generation
von: Rangwani, Harsh, et al.
Veröffentlicht: (2024)
von: Rangwani, Harsh, et al.
Veröffentlicht: (2024)
MoGen: A Unified Collaborative Framework for Controllable Multi-Object Image Generation
von: Li, Yanfeng, et al.
Veröffentlicht: (2026)
von: Li, Yanfeng, et al.
Veröffentlicht: (2026)
AnyControl: Create Your Artwork with Versatile Control on Text-to-Image Generation
von: Sun, Yanan, et al.
Veröffentlicht: (2024)
von: Sun, Yanan, et al.
Veröffentlicht: (2024)
SceneDesigner: Controllable Multi-Object Image Generation with 9-DoF Pose Manipulation
von: Qin, Zhenyuan, et al.
Veröffentlicht: (2025)
von: Qin, Zhenyuan, et al.
Veröffentlicht: (2025)
TOUCH: Text-guided Controllable Generation of Free-Form Hand-Object Interactions
von: Han, Guangyi, et al.
Veröffentlicht: (2025)
von: Han, Guangyi, et al.
Veröffentlicht: (2025)
Controllable 3D Object Generation with Single Image Prompt
von: Lee, Jaeseok, et al.
Veröffentlicht: (2025)
von: Lee, Jaeseok, et al.
Veröffentlicht: (2025)
SINGAPO: Single Image Controlled Generation of Articulated Parts in Objects
von: Liu, Jiayi, et al.
Veröffentlicht: (2024)
von: Liu, Jiayi, et al.
Veröffentlicht: (2024)
MirrorVerse: Pushing Diffusion Models to Realistically Reflect the World
von: Dhiman, Ankit, et al.
Veröffentlicht: (2025)
von: Dhiman, Ankit, et al.
Veröffentlicht: (2025)
Text-guided Controllable Diffusion for Realistic Camouflage Images Generation
von: Qian, Yuhang, et al.
Veröffentlicht: (2025)
von: Qian, Yuhang, et al.
Veröffentlicht: (2025)
Dynamic Frequency Modulation for Controllable Text-driven Image Generation
von: Shi, Tiandong, et al.
Veröffentlicht: (2026)
von: Shi, Tiandong, et al.
Veröffentlicht: (2026)
Controllable Generation with Text-to-Image Diffusion Models: A Survey
von: Cao, Pu, et al.
Veröffentlicht: (2024)
von: Cao, Pu, et al.
Veröffentlicht: (2024)
T2AV-Compass: Towards Unified Evaluation for Text-to-Audio-Video Generation
von: Cao, Zhe, et al.
Veröffentlicht: (2025)
von: Cao, Zhe, et al.
Veröffentlicht: (2025)
Cocktail: Mixing Multi-Modality Controls for Text-Conditional Image Generation
von: Hu, Minghui, et al.
Veröffentlicht: (2023)
von: Hu, Minghui, et al.
Veröffentlicht: (2023)
CREA: A Collaborative Multi-Agent Framework for Creative Image Editing and Generation
von: Venkatesh, Kavana, et al.
Veröffentlicht: (2025)
von: Venkatesh, Kavana, et al.
Veröffentlicht: (2025)
RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation
von: Pang, Lexi, et al.
Veröffentlicht: (2025)
von: Pang, Lexi, et al.
Veröffentlicht: (2025)
GOOD: Towards Domain Generalized Orientated Object Detection
von: Bi, Qi, et al.
Veröffentlicht: (2024)
von: Bi, Qi, et al.
Veröffentlicht: (2024)
FlexEControl: Flexible and Efficient Multimodal Control for Text-to-Image Generation
von: He, Xuehai, et al.
Veröffentlicht: (2024)
von: He, Xuehai, et al.
Veröffentlicht: (2024)
Camera Control for Text-to-Image Generation via Learning Viewpoint Tokens
von: Lu, Xinxuan, et al.
Veröffentlicht: (2026)
von: Lu, Xinxuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PreciseControl: Enhancing Text-To-Image Diffusion Models with Fine-Grained Attribute Control
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024) -
Text2Place: Affordance-aware Text Guided Human Placement
von: Parihar, Rishubh, et al.
Veröffentlicht: (2024) -
SeeThrough3D: Occlusion Aware 3D Control in Text-to-Image Generation
von: Agrawal, Vaibhav, et al.
Veröffentlicht: (2026) -
MonoPlace3D: Learning 3D-Aware Object Placement for 3D Monocular Detection
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025) -
Kontinuous Kontext: Continuous Strength Control for Instruction-based Image Editing
von: Parihar, Rishubh, et al.
Veröffentlicht: (2025)