Adaptive Prompt Elicitation for Text-to-Image Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wen, Xinyi, Hegemann, Lena, Jin, Xiaofu, Ma, Shuai, Oulasvirta, Antti |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RClicks: Realistic Click Simulation for Benchmarking Interactive Segmentation
von: Antonov, Anton, et al.
Veröffentlicht: (2024)
von: Antonov, Anton, et al.
Veröffentlicht: (2024)
An Image-based Typology for Visualization
von: Chen, Jian, et al.
Veröffentlicht: (2024)
von: Chen, Jian, et al.
Veröffentlicht: (2024)
AIM 2024 Challenge on Video Saliency Prediction: Methods and Results
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2024)
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2024)
NTIRE 2026 Challenge on Video Saliency Prediction: Methods and Results
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2026)
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2026)
TouchInsight: Uncertainty-aware Rapid Touch and Text Input for Mixed Reality from Egocentric Vision
von: Streli, Paul, et al.
Veröffentlicht: (2024)
von: Streli, Paul, et al.
Veröffentlicht: (2024)
BREPS: Bounding-Box Robustness Evaluation of Promptable Segmentation
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2026)
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2026)
Explaining Explainability: Recommendations for Effective Use of Concept Activation Vectors
von: Nicolson, Angus, et al.
Veröffentlicht: (2024)
von: Nicolson, Angus, et al.
Veröffentlicht: (2024)
TETRIS: Towards Exploring the Robustness of Interactive Segmentation
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2024)
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2024)
Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark
von: Nagy, Rajmund, et al.
Veröffentlicht: (2025)
von: Nagy, Rajmund, et al.
Veröffentlicht: (2025)
Interactive Groupwise Comparison for Reinforcement Learning from Human Feedback
von: Kompatscher, Jan, et al.
Veröffentlicht: (2025)
von: Kompatscher, Jan, et al.
Veröffentlicht: (2025)
Patient-Specific Dynamic Digital-Physical Twin for Coronary Intervention Training: An Integrated Mixed Reality Approach
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
von: Wang, Shuo, et al.
Veröffentlicht: (2025)
ControlGUI: Guiding Generative GUI Exploration through Perceptual Visual Flow
von: Garg, Aryan, et al.
Veröffentlicht: (2025)
von: Garg, Aryan, et al.
Veröffentlicht: (2025)
RealTalk: Realistic Emotion-Aware Lifelike Talking-Head Synthesis
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
von: Wang, Wenqing, et al.
Veröffentlicht: (2025)
AI Assisted AR Assembly: Object Recognition and Computer Vision for Augmented Reality Assisted Assembly
von: Kyaw, Alexander Htet, et al.
Veröffentlicht: (2025)
von: Kyaw, Alexander Htet, et al.
Veröffentlicht: (2025)
Automatic Dance Video Segmentation for Understanding Choreography
von: Endo, Koki, et al.
Veröffentlicht: (2024)
von: Endo, Koki, et al.
Veröffentlicht: (2024)
Pinching Visuo-haptic Display: Investigating Cross-Modal Effects of Visual Textures on Electrostatic Cloth Tactile Sensations
von: Kitagishi, Takekazu, et al.
Veröffentlicht: (2025)
von: Kitagishi, Takekazu, et al.
Veröffentlicht: (2025)
Real Time Captioning of Sign Language Gestures in Video Meetings
von: Mukherjee, Sharanya, et al.
Veröffentlicht: (2025)
von: Mukherjee, Sharanya, et al.
Veröffentlicht: (2025)
Claycode: Stylable and Deformable 2D Scannable Codes
von: Maida, Marco, et al.
Veröffentlicht: (2025)
von: Maida, Marco, et al.
Veröffentlicht: (2025)
EEG-D3: A Solution to the Hidden Overfitting Problem of Deep Learning Models
von: Ludwig, Siegfried, et al.
Veröffentlicht: (2025)
von: Ludwig, Siegfried, et al.
Veröffentlicht: (2025)
Visual Aesthetic Benchmark: Can Frontier Models Judge Beauty?
von: Feng, Yichen, et al.
Veröffentlicht: (2026)
von: Feng, Yichen, et al.
Veröffentlicht: (2026)
Talking Tennis: Language Feedback from 3D Biomechanical Action Recognition
von: Dashore, Arushi, et al.
Veröffentlicht: (2025)
von: Dashore, Arushi, et al.
Veröffentlicht: (2025)
Towards AI-Architecture Liberty: A Comprehensive Survey on Design and Generation of Virtual Architecture by Deep Learning
von: Wang, Anqi, et al.
Veröffentlicht: (2023)
von: Wang, Anqi, et al.
Veröffentlicht: (2023)
UPAM: Unified Prompt Attack in Text-to-Image Generation Models Against Both Textual Filters and Visual Checkers
von: Peng, Duo, et al.
Veröffentlicht: (2024)
von: Peng, Duo, et al.
Veröffentlicht: (2024)
WheelPoser: Sparse-IMU Based Body Pose Estimation for Wheelchair Users
von: Li, Yunzhi, et al.
Veröffentlicht: (2024)
von: Li, Yunzhi, et al.
Veröffentlicht: (2024)
Analyzing Quality, Bias, and Performance in Text-to-Image Generative Models
von: Masrourisaadat, Nila, et al.
Veröffentlicht: (2024)
von: Masrourisaadat, Nila, et al.
Veröffentlicht: (2024)
DexAvatar: 3D Sign Language Reconstruction with Hand and Body Pose Priors
von: Kundu, Kaustubh, et al.
Veröffentlicht: (2025)
von: Kundu, Kaustubh, et al.
Veröffentlicht: (2025)
A Medical Low-Back Pain Physical Rehabilitation Dataset for Human Body Movement Analysis
von: Nguyen, Sao Mai, et al.
Veröffentlicht: (2024)
von: Nguyen, Sao Mai, et al.
Veröffentlicht: (2024)
DISCOVER: Identifying Patterns of Daily Living in Human Activities from Smart Home Data
von: Karpekov, Alexander, et al.
Veröffentlicht: (2025)
von: Karpekov, Alexander, et al.
Veröffentlicht: (2025)
Towards a GENEA Leaderboard -- an Extended, Living Benchmark for Evaluating and Advancing Conversational Motion Synthesis
von: Nagy, Rajmund, et al.
Veröffentlicht: (2024)
von: Nagy, Rajmund, et al.
Veröffentlicht: (2024)
Topological Structure Description for Artcode Detection Using the Shape of Orientation Histogram
von: Xu, Liming, et al.
Veröffentlicht: (2025)
von: Xu, Liming, et al.
Veröffentlicht: (2025)
TextCAVs: Debugging vision models using text
von: Nicolson, Angus, et al.
Veröffentlicht: (2024)
von: Nicolson, Angus, et al.
Veröffentlicht: (2024)
"Does the cafe entrance look accessible? Where is the door?" Towards Geospatial AI Agents for Visual Inquiries
von: Froehlich, Jon E., et al.
Veröffentlicht: (2025)
von: Froehlich, Jon E., et al.
Veröffentlicht: (2025)
No Longer Trending on Artstation: Prompt Analysis of Generative AI Art
von: McCormack, Jon, et al.
Veröffentlicht: (2024)
von: McCormack, Jon, et al.
Veröffentlicht: (2024)
Humanizing AI Grading: Student-Centered Insights on Fairness, Trust, Consistency and Transparency
von: Riahi, Bahare, et al.
Veröffentlicht: (2026)
von: Riahi, Bahare, et al.
Veröffentlicht: (2026)
Graph4GUI: Graph Neural Networks for Representing Graphical User Interfaces
von: Jiang, Yue, et al.
Veröffentlicht: (2024)
von: Jiang, Yue, et al.
Veröffentlicht: (2024)
SIGGesture: Generalized Co-Speech Gesture Synthesis via Semantic Injection with Large-Scale Pre-Training Diffusion Models
von: Cheng, Qingrong, et al.
Veröffentlicht: (2024)
von: Cheng, Qingrong, et al.
Veröffentlicht: (2024)
Neurocognitive Modeling for Text Generation: Deep Learning Architecture for EEG Data
von: Khushiyant
Veröffentlicht: (2025)
von: Khushiyant
Veröffentlicht: (2025)
Multimodal Fusion of EMG and Vision for Human Grasp Intent Inference in Prosthetic Hand Control
von: Zandigohar, Mehrshad, et al.
Veröffentlicht: (2021)
von: Zandigohar, Mehrshad, et al.
Veröffentlicht: (2021)
Enhancing Cross-Modal Contextual Congruence for Crowdfunding Success using Knowledge-infused Learning
von: Padhi, Trilok, et al.
Veröffentlicht: (2024)
von: Padhi, Trilok, et al.
Veröffentlicht: (2024)
Quantifying the Impact of Motion on 2D Gaze Estimation in Real-World Mobile Interactions
von: Lei, Yaxiong, et al.
Veröffentlicht: (2025)
von: Lei, Yaxiong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RClicks: Realistic Click Simulation for Benchmarking Interactive Segmentation
von: Antonov, Anton, et al.
Veröffentlicht: (2024) -
An Image-based Typology for Visualization
von: Chen, Jian, et al.
Veröffentlicht: (2024) -
AIM 2024 Challenge on Video Saliency Prediction: Methods and Results
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2024) -
NTIRE 2026 Challenge on Video Saliency Prediction: Methods and Results
von: Moskalenko, Andrey, et al.
Veröffentlicht: (2026) -
TouchInsight: Uncertainty-aware Rapid Touch and Text Input for Mixed Reality from Egocentric Vision
von: Streli, Paul, et al.
Veröffentlicht: (2024)