SketchFlex: Facilitating Spatial-Semantic Coherence in Text-to-Image Generation with Region-Based Sketches
Fuente:
arXiv
Guardado en:
| Autores principales: | Lin, Haichuan, Ye, Yilin, Xia, Jiazhi, Zeng, Wei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
AKRMap: Adaptive Kernel Regression for Trustworthy Visualization of Cross-Modal Embeddings
por: Ye, Yilin, et al.
Publicado: (2025)
por: Ye, Yilin, et al.
Publicado: (2025)
SketchPlay: Intuitive Creation of Physically Realistic VR Content with Gesture-Driven Sketching
por: Zhang, Xiangwen, et al.
Publicado: (2025)
por: Zhang, Xiangwen, et al.
Publicado: (2025)
GroundUp: Rapid Sketch-Based 3D City Massing
por: Unlu, Gizem Esra, et al.
Publicado: (2024)
por: Unlu, Gizem Esra, et al.
Publicado: (2024)
Deep Sketch-Based 3D Modeling: A Survey
por: Tono, Alberto, et al.
Publicado: (2026)
por: Tono, Alberto, et al.
Publicado: (2026)
Precise Workcell Sketching from Point Clouds Using an AR Toolbox
por: Zieliński, Krzysztof, et al.
Publicado: (2024)
por: Zieliński, Krzysztof, et al.
Publicado: (2024)
DAVINCI: A Single-Stage Architecture for Constrained CAD Sketch Inference
por: Karadeniz, Ahmet Serdar, et al.
Publicado: (2024)
por: Karadeniz, Ahmet Serdar, et al.
Publicado: (2024)
MS2Mesh-XR: Multi-modal Sketch-to-Mesh Generation in XR Environments
por: Tong, Yuqi, et al.
Publicado: (2024)
por: Tong, Yuqi, et al.
Publicado: (2024)
Inkspire: Supporting Design Exploration with Generative AI through Analogical Sketching
por: Lin, David Chuan-En, et al.
Publicado: (2025)
por: Lin, David Chuan-En, et al.
Publicado: (2025)
AnyUser: Translating Sketched User Intent into Domestic Robots
por: Yang, Songyuan, et al.
Publicado: (2026)
por: Yang, Songyuan, et al.
Publicado: (2026)
Sketch2Prototype: Rapid Conceptual Design Exploration and Prototyping with Generative AI
por: Edwards, Kristen M., et al.
Publicado: (2024)
por: Edwards, Kristen M., et al.
Publicado: (2024)
OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI Collaboration
por: Duan, Junwen, et al.
Publicado: (2025)
por: Duan, Junwen, et al.
Publicado: (2025)
Sketch2Colab: Sketch-Conditioned Multi-Human Animation via Controllable Flow Distillation
por: Daiya, Divyanshu, et al.
Publicado: (2026)
por: Daiya, Divyanshu, et al.
Publicado: (2026)
Semantic Draw Engineering for Text-to-Image Creation
por: Li, Yang, et al.
Publicado: (2023)
por: Li, Yang, et al.
Publicado: (2023)
SketchDynamics: Exploring Free-Form Sketches for Dynamic Intent Expression in Animation Generation
por: Li, Boyu, et al.
Publicado: (2026)
por: Li, Boyu, et al.
Publicado: (2026)
Sketch Bug: Using Sketch-Based Input for Interactive Code Debugging
por: Chen, Helen Weixu, et al.
Publicado: (2026)
por: Chen, Helen Weixu, et al.
Publicado: (2026)
IntentTuner: An Interactive Framework for Integrating Human Intents in Fine-tuning Text-to-Image Generative Models
por: Zeng, Xingchen, et al.
Publicado: (2024)
por: Zeng, Xingchen, et al.
Publicado: (2024)
SketchConcept: Sketching-based Concept Recomposition for Product Design using Generative AI
por: Duan, Runlin, et al.
Publicado: (2025)
por: Duan, Runlin, et al.
Publicado: (2025)
TalkSketch: Multimodal Generative AI for Real-time Sketch Ideation with Speech
por: Shi, Weiyan, et al.
Publicado: (2025)
por: Shi, Weiyan, et al.
Publicado: (2025)
ModalChorus: Visual Probing and Alignment of Multi-modal Embeddings via Modal Fusion Map
por: Ye, Yilin, et al.
Publicado: (2024)
por: Ye, Yilin, et al.
Publicado: (2024)
POET: Supporting Prompting Creativity and Personalization with Automated Expansion of Text-to-Image Generation
por: Han, Evans Xu, et al.
Publicado: (2025)
por: Han, Evans Xu, et al.
Publicado: (2025)
Sketch Then Generate: Providing Incremental User Feedback and Guiding LLM Code Generation through Language-Oriented Code Sketches
por: Zhu-Tian, Chen, et al.
Publicado: (2024)
por: Zhu-Tian, Chen, et al.
Publicado: (2024)
HolmeSketcher: Generative 3D Sketch Mapping for Spatial Reconstruction in Crime Scene Investigation
por: Xiao, Tianyi, et al.
Publicado: (2026)
por: Xiao, Tianyi, et al.
Publicado: (2026)
ImaGGen: Zero-Shot Generation of Co-Speech Semantic Gestures Grounded in Language and Image Input
por: Voss, Hendric, et al.
Publicado: (2025)
por: Voss, Hendric, et al.
Publicado: (2025)
GenColor: Generative Color-Concept Association in Visual Design
por: Hou, Yihan, et al.
Publicado: (2025)
por: Hou, Yihan, et al.
Publicado: (2025)
Text-to-Image Representativity Fairness Evaluation Framework
por: Yamani, Asma, et al.
Publicado: (2024)
por: Yamani, Asma, et al.
Publicado: (2024)
Automated Image-Based Identification and Consistent Classification of Fire Patterns with Quantitative Shape Analysis and Spatial Location Identification
por: Liu, Pengkun, et al.
Publicado: (2024)
por: Liu, Pengkun, et al.
Publicado: (2024)
Confidence Contours: Uncertainty-Aware Annotation for Medical Semantic Segmentation
por: Ye, Andre, et al.
Publicado: (2023)
por: Ye, Andre, et al.
Publicado: (2023)
EIT-1M: One Million EEG-Image-Text Pairs for Human Visual-textual Recognition and More
por: Zheng, Xu, et al.
Publicado: (2024)
por: Zheng, Xu, et al.
Publicado: (2024)
A Taxonomy of Human--MLLM Interaction in Early-Stage Sketch-Based Design Ideation
por: Shi, Weiyan, et al.
Publicado: (2026)
por: Shi, Weiyan, et al.
Publicado: (2026)
Bridging Text and Image for Artist Style Transfer via Contrastive Learning
por: Liu, Zhi-Song, et al.
Publicado: (2024)
por: Liu, Zhi-Song, et al.
Publicado: (2024)
Algorithmic Ways of Seeing: Using Object Detection to Facilitate Art Exploration
por: Meyer, Louie Søs, et al.
Publicado: (2024)
por: Meyer, Louie Søs, et al.
Publicado: (2024)
AR-Facilitated Safety Inspection and Fall Hazard Detection on Construction Sites
por: Liu, Jiazhou, et al.
Publicado: (2024)
por: Liu, Jiazhou, et al.
Publicado: (2024)
Using Text-to-Image Generation for Architectural Design Ideation
por: Paananen, Ville, et al.
Publicado: (2023)
por: Paananen, Ville, et al.
Publicado: (2023)
ConceptFactory: Facilitate 3D Object Knowledge Annotation with Object Conceptualization
por: Sun, Jianhua, et al.
Publicado: (2024)
por: Sun, Jianhua, et al.
Publicado: (2024)
Conveying Meaning through Gestures: An Investigation into Semantic Co-Speech Gesture Generation
por: Voss, Hendric, et al.
Publicado: (2025)
por: Voss, Hendric, et al.
Publicado: (2025)
Semantic and Expressive Variation in Image Captions Across Languages
por: Ye, Andre, et al.
Publicado: (2023)
por: Ye, Andre, et al.
Publicado: (2023)
Steering Generative Models for Accessibility: EasyRead Image Generation
por: Dickenmann, Nicolas, et al.
Publicado: (2026)
por: Dickenmann, Nicolas, et al.
Publicado: (2026)
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
por: Yang, Boyin, et al.
Publicado: (2025)
por: Yang, Boyin, et al.
Publicado: (2025)
SemLayer: Semantic-aware Generative Segmentation and Layer Construction for Abstract Icons
por: Xu, Haiyang, et al.
Publicado: (2026)
por: Xu, Haiyang, et al.
Publicado: (2026)
Facilitating Video Story Interaction with Multi-Agent Collaborative System
por: Zhang, Yiwen, et al.
Publicado: (2025)
por: Zhang, Yiwen, et al.
Publicado: (2025)
Ejemplares similares
-
AKRMap: Adaptive Kernel Regression for Trustworthy Visualization of Cross-Modal Embeddings
por: Ye, Yilin, et al.
Publicado: (2025) -
SketchPlay: Intuitive Creation of Physically Realistic VR Content with Gesture-Driven Sketching
por: Zhang, Xiangwen, et al.
Publicado: (2025) -
GroundUp: Rapid Sketch-Based 3D City Massing
por: Unlu, Gizem Esra, et al.
Publicado: (2024) -
Deep Sketch-Based 3D Modeling: A Survey
por: Tono, Alberto, et al.
Publicado: (2026) -
Precise Workcell Sketching from Point Clouds Using an AR Toolbox
por: Zieliński, Krzysztof, et al.
Publicado: (2024)