Demystifying the Potential of ChatGPT-4 Vision for Construction Progress Monitoring
Fuente:
arXiv
Saved in:
| Main Author: | Ersoz, Ahmet Bahaddin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IQAGPT: Image Quality Assessment with Vision-language and ChatGPT Models
by: Chen, Zhihao, et al.
Published: (2023)
by: Chen, Zhihao, et al.
Published: (2023)
Is ChatGPT-5 Ready for Mammogram VQA?
by: Li, Qiang, et al.
Published: (2025)
by: Li, Qiang, et al.
Published: (2025)
GPTDrawer: Enhancing Visual Synthesis through ChatGPT
by: Li, Kun, et al.
Published: (2024)
by: Li, Kun, et al.
Published: (2024)
12 Angry AI Agents: Evaluating Multi-Agent LLM Decision-Making Through Cinematic Jury Deliberation
by: Ersoz, Ahmet Bahaddin
Published: (2026)
by: Ersoz, Ahmet Bahaddin
Published: (2026)
Prompt fidelity of ChatGPT4o / Dall-E3 text-to-image visualisations
by: Spennemann, Dirk HR
Published: (2025)
by: Spennemann, Dirk HR
Published: (2025)
PhD: A ChatGPT-Prompted Visual hallucination Evaluation Dataset
by: Liu, Jiazhen, et al.
Published: (2024)
by: Liu, Jiazhen, et al.
Published: (2024)
Leveraging ChatGPT's Multimodal Vision Capabilities to Rank Satellite Images by Poverty Level: Advancing Tools for Social Science Research
by: Sarmadi, Hamid, et al.
Published: (2025)
by: Sarmadi, Hamid, et al.
Published: (2025)
ChatGPT and biometrics: an assessment of face recognition, gender detection, and age estimation capabilities
by: Hassanpour, Ahmad, et al.
Published: (2024)
by: Hassanpour, Ahmad, et al.
Published: (2024)
Visual Reasoning Evaluation of Grok, Deepseek Janus, Gemini, Qwen, Mistral, and ChatGPT
by: Jegham, Nidhal, et al.
Published: (2025)
by: Jegham, Nidhal, et al.
Published: (2025)
Evaluating ChatGPT's Performance in Classifying Pneumonia from Chest X-Ray Images
by: Prahallad, Pragna, et al.
Published: (2025)
by: Prahallad, Pragna, et al.
Published: (2025)
Intelligent Director: An Automatic Framework for Dynamic Visual Composition using ChatGPT
by: Zheng, Sixiao, et al.
Published: (2024)
by: Zheng, Sixiao, et al.
Published: (2024)
Can ChatGPT Perform Image Splicing Detection? A Preliminary Study
by: Nath, Souradip
Published: (2025)
by: Nath, Souradip
Published: (2025)
How Good is ChatGPT at Audiovisual Deepfake Detection: A Comparative Study of ChatGPT, AI Models and Human Perception
by: Shahzad, Sahibzada Adil, et al.
Published: (2024)
by: Shahzad, Sahibzada Adil, et al.
Published: (2024)
AI-Generated Content Enhanced Computer-Aided Diagnosis Model for Thyroid Nodules: A ChatGPT-Style Assistant
by: Yao, Jincao, et al.
Published: (2024)
by: Yao, Jincao, et al.
Published: (2024)
Assessing Greenspace Attractiveness with ChatGPT, Claude, and Gemini: Do AI Models Reflect Human Perceptions?
by: Malekzadeh, Milad, et al.
Published: (2025)
by: Malekzadeh, Milad, et al.
Published: (2025)
ChatGPT Encounters Morphing Attack Detection: Zero-Shot MAD with Multi-Modal Large Language Models and General Vision Models
by: Zhang, Haoyu, et al.
Published: (2025)
by: Zhang, Haoyu, et al.
Published: (2025)
Demystifying KAN for Vision Tasks: The RepKAN Approach
by: Cheon, Minjong
Published: (2026)
by: Cheon, Minjong
Published: (2026)
Can ChatGPT Learn My Life From a Week of First-Person Video?
by: Harris, Keegan
Published: (2025)
by: Harris, Keegan
Published: (2025)
Actional Atomic-Concept Learning for Demystifying Vision-Language Navigation
by: Lin, Bingqian, et al.
Published: (2023)
by: Lin, Bingqian, et al.
Published: (2023)
How Good is ChatGPT at Face Biometrics? A First Look into Recognition, Soft Biometrics, and Explainability
by: DeAndres-Tame, Ivan, et al.
Published: (2024)
by: DeAndres-Tame, Ivan, et al.
Published: (2024)
Demystifying Video Reasoning
by: Wang, Ruisi, et al.
Published: (2026)
by: Wang, Ruisi, et al.
Published: (2026)
Remote Sensing ChatGPT: Solving Remote Sensing Tasks with ChatGPT and Visual Models
by: Guo, Haonan, et al.
Published: (2024)
by: Guo, Haonan, et al.
Published: (2024)
Demystifying Foreground-Background Memorization in Diffusion Models
by: Di, Jimmy Z., et al.
Published: (2025)
by: Di, Jimmy Z., et al.
Published: (2025)
Evaluation of GPT-4o and GPT-4o-mini's Vision Capabilities for Compositional Analysis from Dried Solution Drops
by: Dangi, Deven B., et al.
Published: (2024)
by: Dangi, Deven B., et al.
Published: (2024)
ChatGPT Meets Iris Biometrics
by: Farmanifard, Parisa, et al.
Published: (2024)
by: Farmanifard, Parisa, et al.
Published: (2024)
GPT as Psychologist? Preliminary Evaluations for GPT-4V on Visual Affective Computing
by: Lu, Hao, et al.
Published: (2024)
by: Lu, Hao, et al.
Published: (2024)
Demystifying the Visual Quality Paradox in Multimodal Large Language Models
by: Xing, Shuo, et al.
Published: (2025)
by: Xing, Shuo, et al.
Published: (2025)
Leveraging Vision-Language Models to Select Trustworthy Super-Resolution Samples Generated by Diffusion Models
by: Korkmaz, Cansu, et al.
Published: (2025)
by: Korkmaz, Cansu, et al.
Published: (2025)
AgroGPT: Efficient Agricultural Vision-Language Model with Expert Tuning
by: Awais, Muhammad, et al.
Published: (2024)
by: Awais, Muhammad, et al.
Published: (2024)
MiniGPT-Reverse-Designing: Predicting Image Adjustments Utilizing MiniGPT-4
by: Azizi, Vahid, et al.
Published: (2024)
by: Azizi, Vahid, et al.
Published: (2024)
Demystifying Visual Features of Movie Posters for Multi-Label Genre Identification
by: Nareti, Utsav Kumar, et al.
Published: (2023)
by: Nareti, Utsav Kumar, et al.
Published: (2023)
Leveraging Chat-Based Large Vision Language Models for Multimodal Out-Of-Context Detection
by: Shalabi, Fatma, et al.
Published: (2024)
by: Shalabi, Fatma, et al.
Published: (2024)
Advancing Complex Video Object Segmentation via Progressive Concept Construction
by: Zhang, Zhixiong, et al.
Published: (2025)
by: Zhang, Zhixiong, et al.
Published: (2025)
Tokensome: Towards a Genetic Vision-Language GPT for Explainable and Cognitive Karyotyping
by: Zhang, Haoxi, et al.
Published: (2024)
by: Zhang, Haoxi, et al.
Published: (2024)
MiniGPT-5: Interleaved Vision-and-Language Generation via Generative Vokens
by: Zheng, Kaizhi, et al.
Published: (2023)
by: Zheng, Kaizhi, et al.
Published: (2023)
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding
by: Chen, Ying, et al.
Published: (2024)
by: Chen, Ying, et al.
Published: (2024)
Automatic Real-time Vehicle Classification by Image Colour Component Based Template Matching
by: Orun, Ahmet
Published: (2022)
by: Orun, Ahmet
Published: (2022)
An Evaluation of GPT-4V and Gemini in Online VQA
by: Liu, Mengchen, et al.
Published: (2023)
by: Liu, Mengchen, et al.
Published: (2023)
UPDP: A Unified Progressive Depth Pruner for CNN and Vision Transformer
by: Liu, Ji, et al.
Published: (2024)
by: Liu, Ji, et al.
Published: (2024)
Vision-Language Models in Remote Sensing: Current Progress and Future Trends
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
Similar Items
-
IQAGPT: Image Quality Assessment with Vision-language and ChatGPT Models
by: Chen, Zhihao, et al.
Published: (2023) -
Is ChatGPT-5 Ready for Mammogram VQA?
by: Li, Qiang, et al.
Published: (2025) -
GPTDrawer: Enhancing Visual Synthesis through ChatGPT
by: Li, Kun, et al.
Published: (2024) -
12 Angry AI Agents: Evaluating Multi-Agent LLM Decision-Making Through Cinematic Jury Deliberation
by: Ersoz, Ahmet Bahaddin
Published: (2026) -
Prompt fidelity of ChatGPT4o / Dall-E3 text-to-image visualisations
by: Spennemann, Dirk HR
Published: (2025)