Multimodal AI for Body Fat Estimation: Computer Vision and Anthropometry with DEXA Benchmarks
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Aldajani, Rayan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks
von: Ramachandran, Rahul, et al.
Veröffentlicht: (2025)
von: Ramachandran, Rahul, et al.
Veröffentlicht: (2025)
FM-G-CAM: A Holistic Approach for Explainable AI in Computer Vision
von: Silva, Ravidu Suien Rammuni, et al.
Veröffentlicht: (2023)
von: Silva, Ravidu Suien Rammuni, et al.
Veröffentlicht: (2023)
DetoxAI: a Python Toolkit for Debiasing Deep Learning Models in Computer Vision
von: Stępka, Ignacy, et al.
Veröffentlicht: (2025)
von: Stępka, Ignacy, et al.
Veröffentlicht: (2025)
Temporal Embeddings: Scalable Self-Supervised Temporal Representation Learning from Spatiotemporal Data for Multimodal Computer Vision
von: Cao, Yi, et al.
Veröffentlicht: (2023)
von: Cao, Yi, et al.
Veröffentlicht: (2023)
Towards Real-Time 2D Mapping: Harnessing Drones, AI, and Computer Vision for Advanced Insights
von: Agnur, Bharath Kumar
Veröffentlicht: (2024)
von: Agnur, Bharath Kumar
Veröffentlicht: (2024)
When Graph meets Multimodal: Benchmarking and Meditating on Multimodal Attributed Graphs Learning
von: Yan, Hao, et al.
Veröffentlicht: (2024)
von: Yan, Hao, et al.
Veröffentlicht: (2024)
Uncertainties of Latent Representations in Computer Vision
von: Kirchhof, Michael
Veröffentlicht: (2024)
von: Kirchhof, Michael
Veröffentlicht: (2024)
Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark
von: Huybrechts, Goeric, et al.
Veröffentlicht: (2025)
von: Huybrechts, Goeric, et al.
Veröffentlicht: (2025)
Vision-DeepResearch Benchmark: Rethinking Visual and Textual Search for Multimodal Large Language Models
von: Zeng, Yu, et al.
Veröffentlicht: (2026)
von: Zeng, Yu, et al.
Veröffentlicht: (2026)
Benchmarking Domain Generalization Algorithms in Computational Pathology
von: Zamanitajeddin, Neda, et al.
Veröffentlicht: (2024)
von: Zamanitajeddin, Neda, et al.
Veröffentlicht: (2024)
SAVER: Selective As-Needed Vision Evidence for Multimodal Information Extraction
von: Hu, Miaobo, et al.
Veröffentlicht: (2026)
von: Hu, Miaobo, et al.
Veröffentlicht: (2026)
IndicVisionBench: Benchmarking Cultural and Multilingual Understanding in VLMs
von: Faraz, Ali, et al.
Veröffentlicht: (2025)
von: Faraz, Ali, et al.
Veröffentlicht: (2025)
A Smart Healthcare System for Monkeypox Skin Lesion Detection and Tracking
von: Alghoraibi, Huda, et al.
Veröffentlicht: (2025)
von: Alghoraibi, Huda, et al.
Veröffentlicht: (2025)
Computer Vision Approaches for Automated Bee Counting Application
von: Bilik, Simon, et al.
Veröffentlicht: (2024)
von: Bilik, Simon, et al.
Veröffentlicht: (2024)
TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation
von: Gong, Han, et al.
Veröffentlicht: (2026)
von: Gong, Han, et al.
Veröffentlicht: (2026)
AVA-Bench: Atomic Visual Ability Benchmark for Vision Foundation Models
von: Mai, Zheda, et al.
Veröffentlicht: (2025)
von: Mai, Zheda, et al.
Veröffentlicht: (2025)
MMErroR: A Benchmark for Erroneous Reasoning in Vision-Language Models
von: Shi, Yang, et al.
Veröffentlicht: (2026)
von: Shi, Yang, et al.
Veröffentlicht: (2026)
Modelling and Simulation of Neuromorphic Datasets for Anomaly Detection in Computer Vision
von: Middleton, Mike, et al.
Veröffentlicht: (2026)
von: Middleton, Mike, et al.
Veröffentlicht: (2026)
MMS-VPR: Multimodal Street-Level Visual Place Recognition Dataset and Benchmark
von: Ou, Yiwei, et al.
Veröffentlicht: (2025)
von: Ou, Yiwei, et al.
Veröffentlicht: (2025)
Circuit Tracing in Vision-Language Models: Understanding the Internal Mechanisms of Multimodal Thinking
von: Yang, Jingcheng, et al.
Veröffentlicht: (2026)
von: Yang, Jingcheng, et al.
Veröffentlicht: (2026)
Benchmarking Feature Upsampling Methods for Vision Foundation Models using Interactive Segmentation
von: Havrylov, Volodymyr, et al.
Veröffentlicht: (2025)
von: Havrylov, Volodymyr, et al.
Veröffentlicht: (2025)
SynthVision -- Harnessing Minimal Input for Maximal Output in Computer Vision Models using Synthetic Image data
von: Kularathne, Yudara, et al.
Veröffentlicht: (2024)
von: Kularathne, Yudara, et al.
Veröffentlicht: (2024)
Unified Supervision For Vision-Language Modeling in 3D Computed Tomography
von: Lee, Hao-Chih, et al.
Veröffentlicht: (2025)
von: Lee, Hao-Chih, et al.
Veröffentlicht: (2025)
On Background Bias of Post-Hoc Concept Embeddings in Computer Vision DNNs
von: Schwalbe, Gesina, et al.
Veröffentlicht: (2025)
von: Schwalbe, Gesina, et al.
Veröffentlicht: (2025)
Generative AI in Vision: A Survey on Models, Metrics and Applications
von: Raut, Gaurav, et al.
Veröffentlicht: (2024)
von: Raut, Gaurav, et al.
Veröffentlicht: (2024)
Collision-Aware Vision-Language Learning for End-to-End Driving with Multimodal Infraction Datasets
von: Koran, Alex, et al.
Veröffentlicht: (2026)
von: Koran, Alex, et al.
Veröffentlicht: (2026)
ThermEval: A Structured Benchmark for Evaluation of Vision-Language Models on Thermal Imagery
von: Shrivastava, Ayush, et al.
Veröffentlicht: (2026)
von: Shrivastava, Ayush, et al.
Veröffentlicht: (2026)
Spatial Atlas: Compute-Grounded Reasoning for Spatial-Aware Research Agent Benchmarks
von: Sharma, Arun
Veröffentlicht: (2026)
von: Sharma, Arun
Veröffentlicht: (2026)
Autonomous AI Surveillance: Multimodal Deep Learning for Cognitive and Behavioral Monitoring
von: Hamza, Ameer, et al.
Veröffentlicht: (2025)
von: Hamza, Ameer, et al.
Veröffentlicht: (2025)
See, Hear, and Understand: Benchmarking Audiovisual Human Speech Understanding in Multimodal Large Language Models
von: Nguyen, Le Thien Phuc, et al.
Veröffentlicht: (2025)
von: Nguyen, Le Thien Phuc, et al.
Veröffentlicht: (2025)
Leaf Angle Estimation using Mask R-CNN and LETR Vision Transformer
von: Margapuri, Venkat, et al.
Veröffentlicht: (2024)
von: Margapuri, Venkat, et al.
Veröffentlicht: (2024)
Back Home: A Computer Vision Solution to Seashell Identification for Ecological Restoration
von: Valverde, Alexander, et al.
Veröffentlicht: (2025)
von: Valverde, Alexander, et al.
Veröffentlicht: (2025)
Caregiver Talk Shapes Toddler Vision: A Computational Study of Dyadic Play
von: Schaumlöffel, Timothy, et al.
Veröffentlicht: (2023)
von: Schaumlöffel, Timothy, et al.
Veröffentlicht: (2023)
Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language Models
von: Ryan, Yuriel, et al.
Veröffentlicht: (2026)
von: Ryan, Yuriel, et al.
Veröffentlicht: (2026)
SPoRC-VIST: A Benchmark for Evaluating Generative Natural Narrative in Vision-Language Models
von: Zeng, Yunlin
Veröffentlicht: (2026)
von: Zeng, Yunlin
Veröffentlicht: (2026)
THRONE: An Object-based Hallucination Benchmark for the Free-form Generations of Large Vision-Language Models
von: Kaul, Prannay, et al.
Veröffentlicht: (2024)
von: Kaul, Prannay, et al.
Veröffentlicht: (2024)
AI-Powered Deepfake Detection Using CNN and Vision Transformer Architectures
von: Urmi, Sifatullah Sheikh, et al.
Veröffentlicht: (2026)
von: Urmi, Sifatullah Sheikh, et al.
Veröffentlicht: (2026)
SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain
von: Mao, Lingtao, et al.
Veröffentlicht: (2026)
von: Mao, Lingtao, et al.
Veröffentlicht: (2026)
When Multi-Task Learning Meets Partial Supervision: A Computer Vision Review
von: Fontana, Maxime, et al.
Veröffentlicht: (2023)
von: Fontana, Maxime, et al.
Veröffentlicht: (2023)
LucidPPN: Unambiguous Prototypical Parts Network for User-centric Interpretable Computer Vision
von: Pach, Mateusz, et al.
Veröffentlicht: (2024)
von: Pach, Mateusz, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks
von: Ramachandran, Rahul, et al.
Veröffentlicht: (2025) -
FM-G-CAM: A Holistic Approach for Explainable AI in Computer Vision
von: Silva, Ravidu Suien Rammuni, et al.
Veröffentlicht: (2023) -
DetoxAI: a Python Toolkit for Debiasing Deep Learning Models in Computer Vision
von: Stępka, Ignacy, et al.
Veröffentlicht: (2025) -
Temporal Embeddings: Scalable Self-Supervised Temporal Representation Learning from Spatiotemporal Data for Multimodal Computer Vision
von: Cao, Yi, et al.
Veröffentlicht: (2023) -
Towards Real-Time 2D Mapping: Harnessing Drones, AI, and Computer Vision for Advanced Insights
von: Agnur, Bharath Kumar
Veröffentlicht: (2024)