Prompt fidelity of ChatGPT4o / Dall-E3 text-to-image visualisations
Fuente:
arXiv
Guardado en:
| Autor principal: | Spennemann, Dirk HR |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PhD: A ChatGPT-Prompted Visual hallucination Evaluation Dataset
por: Liu, Jiazhen, et al.
Publicado: (2024)
por: Liu, Jiazhen, et al.
Publicado: (2024)
Is ChatGPT-5 Ready for Mammogram VQA?
por: Li, Qiang, et al.
Publicado: (2025)
por: Li, Qiang, et al.
Publicado: (2025)
Demystifying the Potential of ChatGPT-4 Vision for Construction Progress Monitoring
por: Ersoz, Ahmet Bahaddin
Publicado: (2024)
por: Ersoz, Ahmet Bahaddin
Publicado: (2024)
GPTDrawer: Enhancing Visual Synthesis through ChatGPT
por: Li, Kun, et al.
Publicado: (2024)
por: Li, Kun, et al.
Publicado: (2024)
IQAGPT: Image Quality Assessment with Vision-language and ChatGPT Models
por: Chen, Zhihao, et al.
Publicado: (2023)
por: Chen, Zhihao, et al.
Publicado: (2023)
Visual Reasoning Evaluation of Grok, Deepseek Janus, Gemini, Qwen, Mistral, and ChatGPT
por: Jegham, Nidhal, et al.
Publicado: (2025)
por: Jegham, Nidhal, et al.
Publicado: (2025)
Evaluating ChatGPT's Performance in Classifying Pneumonia from Chest X-Ray Images
por: Prahallad, Pragna, et al.
Publicado: (2025)
por: Prahallad, Pragna, et al.
Publicado: (2025)
ChatGPT and biometrics: an assessment of face recognition, gender detection, and age estimation capabilities
por: Hassanpour, Ahmad, et al.
Publicado: (2024)
por: Hassanpour, Ahmad, et al.
Publicado: (2024)
Intelligent Director: An Automatic Framework for Dynamic Visual Composition using ChatGPT
por: Zheng, Sixiao, et al.
Publicado: (2024)
por: Zheng, Sixiao, et al.
Publicado: (2024)
"Draw me a curator" Examining the visual stereotyping of a cultural services profession by generative AI
por: Spennemann, Dirk HR
Publicado: (2025)
por: Spennemann, Dirk HR
Publicado: (2025)
3DAxisPrompt: Promoting the 3D Grounding and Reasoning in GPT-4o
por: Liu, Dingning, et al.
Publicado: (2025)
por: Liu, Dingning, et al.
Publicado: (2025)
Can ChatGPT Perform Image Splicing Detection? A Preliminary Study
por: Nath, Souradip
Publicado: (2025)
por: Nath, Souradip
Publicado: (2025)
How Good is ChatGPT at Audiovisual Deepfake Detection: A Comparative Study of ChatGPT, AI Models and Human Perception
por: Shahzad, Sahibzada Adil, et al.
Publicado: (2024)
por: Shahzad, Sahibzada Adil, et al.
Publicado: (2024)
AI-Generated Content Enhanced Computer-Aided Diagnosis Model for Thyroid Nodules: A ChatGPT-Style Assistant
por: Yao, Jincao, et al.
Publicado: (2024)
por: Yao, Jincao, et al.
Publicado: (2024)
Assessing Greenspace Attractiveness with ChatGPT, Claude, and Gemini: Do AI Models Reflect Human Perceptions?
por: Malekzadeh, Milad, et al.
Publicado: (2025)
por: Malekzadeh, Milad, et al.
Publicado: (2025)
Leveraging ChatGPT's Multimodal Vision Capabilities to Rank Satellite Images by Poverty Level: Advancing Tools for Social Science Research
por: Sarmadi, Hamid, et al.
Publicado: (2025)
por: Sarmadi, Hamid, et al.
Publicado: (2025)
Delving into: the quantification of Ai-generated content on the internet (synthetic data)
por: Spennemann, Dirk HR
Publicado: (2025)
por: Spennemann, Dirk HR
Publicado: (2025)
Can ChatGPT Learn My Life From a Week of First-Person Video?
por: Harris, Keegan
Publicado: (2025)
por: Harris, Keegan
Publicado: (2025)
Clinical application of HEDI for biomechanical evaluation and visualisation in incisional hernia repair
por: Lösel, Philipp D., et al.
Publicado: (2023)
por: Lösel, Philipp D., et al.
Publicado: (2023)
How Good is ChatGPT at Face Biometrics? A First Look into Recognition, Soft Biometrics, and Explainability
por: DeAndres-Tame, Ivan, et al.
Publicado: (2024)
por: DeAndres-Tame, Ivan, et al.
Publicado: (2024)
CollagePrompt: A Benchmark for Budget-Friendly Visual Recognition with GPT-4V
por: Xu, Siyu, et al.
Publicado: (2024)
por: Xu, Siyu, et al.
Publicado: (2024)
ChatGPT Encounters Morphing Attack Detection: Zero-Shot MAD with Multi-Modal Large Language Models and General Vision Models
por: Zhang, Haoyu, et al.
Publicado: (2025)
por: Zhang, Haoyu, et al.
Publicado: (2025)
How unique are hallucinated citations offered by generative Artificial Intelligence models?
por: Spennemann, Dirk HR
Publicado: (2026)
por: Spennemann, Dirk HR
Publicado: (2026)
Remote Sensing ChatGPT: Solving Remote Sensing Tasks with ChatGPT and Visual Models
por: Guo, Haonan, et al.
Publicado: (2024)
por: Guo, Haonan, et al.
Publicado: (2024)
ChatGPT Meets Iris Biometrics
por: Farmanifard, Parisa, et al.
Publicado: (2024)
por: Farmanifard, Parisa, et al.
Publicado: (2024)
Evaluation of GPT-4o and GPT-4o-mini's Vision Capabilities for Compositional Analysis from Dried Solution Drops
por: Dangi, Deven B., et al.
Publicado: (2024)
por: Dangi, Deven B., et al.
Publicado: (2024)
Comparative Analysis of OpenAI GPT-4o and DeepSeek R1 for Scientific Text Categorization Using Prompt Engineering
por: Maiti, Aniruddha, et al.
Publicado: (2025)
por: Maiti, Aniruddha, et al.
Publicado: (2025)
GPT as Psychologist? Preliminary Evaluations for GPT-4V on Visual Affective Computing
por: Lu, Hao, et al.
Publicado: (2024)
por: Lu, Hao, et al.
Publicado: (2024)
MCAD: Multi-teacher Cross-modal Alignment Distillation for efficient image-text retrieval
por: Lei, Youbo, et al.
Publicado: (2023)
por: Lei, Youbo, et al.
Publicado: (2023)
MiniGPT-Reverse-Designing: Predicting Image Adjustments Utilizing MiniGPT-4
por: Azizi, Vahid, et al.
Publicado: (2024)
por: Azizi, Vahid, et al.
Publicado: (2024)
ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation
por: Chen, Junying, et al.
Publicado: (2025)
por: Chen, Junying, et al.
Publicado: (2025)
Towards High-fidelity 3D Talking Avatar with Personalized Dynamic Texture
por: Li, Xuanchen, et al.
Publicado: (2025)
por: Li, Xuanchen, et al.
Publicado: (2025)
OpenGPT-4o-Image: A Comprehensive Dataset for Advanced Image Generation and Editing
por: Chen, Zhihong, et al.
Publicado: (2025)
por: Chen, Zhihong, et al.
Publicado: (2025)
TV-SAM: Increasing Zero-Shot Segmentation Performance on Multimodal Medical Images Using GPT-4 Generated Descriptive Prompts Without Human Annotation
por: Jiang, Zekun, et al.
Publicado: (2024)
por: Jiang, Zekun, et al.
Publicado: (2024)
A comprehensive study on fidelity metrics for XAI
por: Miró-Nicolau, Miquel, et al.
Publicado: (2024)
por: Miró-Nicolau, Miquel, et al.
Publicado: (2024)
Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation
por: Ye, Junyan, et al.
Publicado: (2025)
por: Ye, Junyan, et al.
Publicado: (2025)
An Evaluation of GPT-4V and Gemini in Online VQA
por: Liu, Mengchen, et al.
Publicado: (2023)
por: Liu, Mengchen, et al.
Publicado: (2023)
Automated Prompt Generation for Creative and Counterfactual Text-to-image Synthesis
por: Jelaca, Aleksa, et al.
Publicado: (2025)
por: Jelaca, Aleksa, et al.
Publicado: (2025)
High-fidelity Person-centric Subject-to-Image Synthesis
por: Wang, Yibin, et al.
Publicado: (2023)
por: Wang, Yibin, et al.
Publicado: (2023)
Will GPT-4 Run DOOM?
por: de Wynter, Adrian
Publicado: (2024)
por: de Wynter, Adrian
Publicado: (2024)
Ejemplares similares
-
PhD: A ChatGPT-Prompted Visual hallucination Evaluation Dataset
por: Liu, Jiazhen, et al.
Publicado: (2024) -
Is ChatGPT-5 Ready for Mammogram VQA?
por: Li, Qiang, et al.
Publicado: (2025) -
Demystifying the Potential of ChatGPT-4 Vision for Construction Progress Monitoring
por: Ersoz, Ahmet Bahaddin
Publicado: (2024) -
GPTDrawer: Enhancing Visual Synthesis through ChatGPT
por: Li, Kun, et al.
Publicado: (2024) -
IQAGPT: Image Quality Assessment with Vision-language and ChatGPT Models
por: Chen, Zhihao, et al.
Publicado: (2023)