Towards In-Vehicle Multi-Task Facial Attribute Recognition: Investigating Synthetic Data and Vision Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Seraj, Esmaeil, Talamonti, Walter |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards a Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition
by: Ramirez, David F., et al.
Published: (2026)
by: Ramirez, David F., et al.
Published: (2026)
Are Vision Foundation Models Ready for Out-of-the-Box Medical Image Registration?
by: Gu, Hanxue, et al.
Published: (2025)
by: Gu, Hanxue, et al.
Published: (2025)
Foundation Artificial Intelligence Models for Health Recognition Using Face Photographs (FAHR-Face)
by: Haugg, Fridolin, et al.
Published: (2025)
by: Haugg, Fridolin, et al.
Published: (2025)
A Survey of Multimodal Ophthalmic Diagnostics: From Task-Specific Approaches to Foundational Models
by: Luo, Xiaoling, et al.
Published: (2025)
by: Luo, Xiaoling, et al.
Published: (2025)
Towards a Multimodal MRI-Based Foundation Model for Multi-Level Feature Exploration in Segmentation, Molecular Subtyping, and Grading of Glioma
by: Farahani, Somayeh, et al.
Published: (2025)
by: Farahani, Somayeh, et al.
Published: (2025)
A Disease-Specific Foundation Model Using Over 100K Fundus Images: Release and Validation for Abnormality and Multi-Disease Classification on Downstream Tasks
by: Jang, Boa, et al.
Published: (2024)
by: Jang, Boa, et al.
Published: (2024)
A Disease-Centric Vision-Language Foundation Model for Precision Oncology in Kidney Cancer
by: Tao, Yuhui, et al.
Published: (2025)
by: Tao, Yuhui, et al.
Published: (2025)
MetaFruit Meets Foundation Models: Leveraging a Comprehensive Multi-Fruit Dataset for Advancing Agricultural Foundation Models
by: Li, Jiajia, et al.
Published: (2024)
by: Li, Jiajia, et al.
Published: (2024)
An Empirical Study on the Fairness of Foundation Models for Multi-Organ Image Segmentation
by: Li, Qin, et al.
Published: (2024)
by: Li, Qin, et al.
Published: (2024)
Multimodal, Multi-Disease Medical Imaging Foundation Model (MerMED-FM)
by: Zhou, Yang, et al.
Published: (2025)
by: Zhou, Yang, et al.
Published: (2025)
When Do Domain-Specific Foundation Models Justify Their Cost? A Systematic Evaluation Across Retinal Imaging Tasks
by: Isztl, David, et al.
Published: (2025)
by: Isztl, David, et al.
Published: (2025)
Self-Supervised Backbone Framework for Diverse Agricultural Vision Tasks
by: Sornapudi, Sudhir, et al.
Published: (2024)
by: Sornapudi, Sudhir, et al.
Published: (2024)
Fine-Grained Cat Breed Recognition with Global Context Vision Transformer
by: Hera, Mowmita Parvin, et al.
Published: (2026)
by: Hera, Mowmita Parvin, et al.
Published: (2026)
Towards Cardiac MRI Foundation Models: Comprehensive Visual-Tabular Representations for Whole-Heart Assessment and Beyond
by: Zhang, Yundi, et al.
Published: (2025)
by: Zhang, Yundi, et al.
Published: (2025)
Synthetic Data in Radiological Imaging: Current State and Future Outlook
by: Sizikova, Elena, et al.
Published: (2024)
by: Sizikova, Elena, et al.
Published: (2024)
Synthetic Data Generation of Body Motion Data by Neural Gas Network for Emotion Recognition
by: Mousavi, Seyed Muhammad Hossein
Published: (2025)
by: Mousavi, Seyed Muhammad Hossein
Published: (2025)
From Spaceborne to Airborne: SAR Image Synthesis Using Foundation Models for Multi-Scale Adaptation
by: Debuysere, Solene, et al.
Published: (2025)
by: Debuysere, Solene, et al.
Published: (2025)
Pathological MRI Segmentation by Synthetic Pathological Data Generation in Fetuses and Neonates
by: Kaandorp, Misha P. T, et al.
Published: (2025)
by: Kaandorp, Misha P. T, et al.
Published: (2025)
Domain-Transferred Synthetic Data Generation for Improving Monocular Depth Estimation
by: Lee, Seungyeop, et al.
Published: (2024)
by: Lee, Seungyeop, et al.
Published: (2024)
Curriculum Learning with Synthetic Data for Enhanced Pulmonary Nodule Detection in Chest Radiographs
by: Sambhu, Pranav, et al.
Published: (2025)
by: Sambhu, Pranav, et al.
Published: (2025)
SynthFM: Training Modality-agnostic Foundation Models for Medical Image Segmentation without Real Medical Data
by: Sengupta, Sourya, et al.
Published: (2025)
by: Sengupta, Sourya, et al.
Published: (2025)
Vision-Language Synthetic Data Enhances Echocardiography Downstream Tasks
by: Ashrafian, Pooria, et al.
Published: (2024)
by: Ashrafian, Pooria, et al.
Published: (2024)
VisionFM: a Multi-Modal Multi-Task Vision Foundation Model for Generalist Ophthalmic Artificial Intelligence
by: Qiu, Jianing, et al.
Published: (2023)
by: Qiu, Jianing, et al.
Published: (2023)
Dual Thinking and Logical Processing -- Are Multi-modal Large Language Models Closing the Gap with Human Vision ?
by: Dayanandan, Kailas, et al.
Published: (2024)
by: Dayanandan, Kailas, et al.
Published: (2024)
Foundation Models for Medical Imaging: Status, Challenges, and Directions
by: Niu, Chuang, et al.
Published: (2026)
by: Niu, Chuang, et al.
Published: (2026)
Foundation Models in Medical Imaging: A Review and Outlook
by: van Veldhuizen, Vivien, et al.
Published: (2025)
by: van Veldhuizen, Vivien, et al.
Published: (2025)
DeepMultiConnectome: Deep Multi-Task Prediction of Structural Connectomes Directly from Diffusion MRI Tractography
by: Vroemen, Marcus J., et al.
Published: (2025)
by: Vroemen, Marcus J., et al.
Published: (2025)
Assessment of Cell Nuclei AI Foundation Models in Kidney Pathology
by: Guo, Junlin, et al.
Published: (2024)
by: Guo, Junlin, et al.
Published: (2024)
orGAN: A Synthetic Data Augmentation Pipeline for Simultaneous Generation of Surgical Images and Ground Truth Labels
by: Nataraj, Niran, et al.
Published: (2025)
by: Nataraj, Niran, et al.
Published: (2025)
VisionUnite: A Vision-Language Foundation Model for Ophthalmology Enhanced with Clinical Knowledge
by: Li, Zihan, et al.
Published: (2024)
by: Li, Zihan, et al.
Published: (2024)
Bladder Cancer Diagnosis with Deep Learning: A Multi-Task Framework and Online Platform
by: Yu, Jinliang, et al.
Published: (2025)
by: Yu, Jinliang, et al.
Published: (2025)
Adversarial Multi-Task Learning for Liver Tumor Segmentation, Dynamic Enhancement Regression, and Classification
by: Xiao, Xiaojiao, et al.
Published: (2025)
by: Xiao, Xiaojiao, et al.
Published: (2025)
ONCOPILOT: A Promptable CT Foundation Model For Solid Tumor Evaluation
by: Machado, Léo, et al.
Published: (2024)
by: Machado, Léo, et al.
Published: (2024)
USF-MAE: Ultrasound Self-Supervised Foundation Model with Masked Autoencoding
by: Megahed, Youssef, et al.
Published: (2025)
by: Megahed, Youssef, et al.
Published: (2025)
Enhancing Diagnostic Reliability of Foundation Model with Uncertainty Estimation in OCT Images
by: Peng, Yuanyuan, et al.
Published: (2024)
by: Peng, Yuanyuan, et al.
Published: (2024)
Improving Representation of High-frequency Components for Medical Visual Foundation Models
by: Chu, Yuetan, et al.
Published: (2024)
by: Chu, Yuetan, et al.
Published: (2024)
MAISI: Medical AI for Synthetic Imaging
by: Guo, Pengfei, et al.
Published: (2024)
by: Guo, Pengfei, et al.
Published: (2024)
Face-MakeUpV2: Facial Consistency Learning for Controllable Text-to-Image Generation
by: Dai, Dawei, et al.
Published: (2025)
by: Dai, Dawei, et al.
Published: (2025)
Towards Blind Bitstream-corrupted Video Recovery via a Visual Foundation Model-driven Framework
by: Liu, Tianyi, et al.
Published: (2025)
by: Liu, Tianyi, et al.
Published: (2025)
From Pretraining to Privacy: Federated Ultrasound Foundation Model with Self-Supervised Learning
by: Jiang, Yuncheng, et al.
Published: (2024)
by: Jiang, Yuncheng, et al.
Published: (2024)
Similar Items
-
Towards a Large Language-Vision Question Answering Model for MSTAR Automatic Target Recognition
by: Ramirez, David F., et al.
Published: (2026) -
Are Vision Foundation Models Ready for Out-of-the-Box Medical Image Registration?
by: Gu, Hanxue, et al.
Published: (2025) -
Foundation Artificial Intelligence Models for Health Recognition Using Face Photographs (FAHR-Face)
by: Haugg, Fridolin, et al.
Published: (2025) -
A Survey of Multimodal Ophthalmic Diagnostics: From Task-Specific Approaches to Foundational Models
by: Luo, Xiaoling, et al.
Published: (2025) -
Towards a Multimodal MRI-Based Foundation Model for Multi-Level Feature Exploration in Segmentation, Molecular Subtyping, and Grading of Glioma
by: Farahani, Somayeh, et al.
Published: (2025)