Multi-modal Machine Learning for Vehicle Rating Predictions Using Image, Text, and Parametric Data
Fuente:
arXiv
Saved in:
| Main Authors: | Su, Hanqi, Song, Binyang, Ahmed, Faez |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BIKED++: A Multimodal Dataset of 1.4 Million Bicycle Image and Parametric CAD Designs
by: Regenwetter, Lyle, et al.
Published: (2024)
by: Regenwetter, Lyle, et al.
Published: (2024)
Simultaneous Long-tailed Recognition and Multi-modal Fusion for Highly Imbalanced Multi-modal Data
by: Yoon, Heegeon, et al.
Published: (2026)
by: Yoon, Heegeon, et al.
Published: (2026)
Multi-modal Vision Pre-training for Medical Image Analysis
by: Rui, Shaohao, et al.
Published: (2024)
by: Rui, Shaohao, et al.
Published: (2024)
Integrating Text and Image Pre-training for Multi-modal Algorithmic Reasoning
by: Zhang, Zijian, et al.
Published: (2024)
by: Zhang, Zijian, et al.
Published: (2024)
Robust Disaster Assessment from Aerial Imagery Using Text-to-Image Synthetic Data
by: Kalluri, Tarun, et al.
Published: (2024)
by: Kalluri, Tarun, et al.
Published: (2024)
SurvMamba: State Space Model with Multi-grained Multi-modal Interaction for Survival Prediction
by: Chen, Ying, et al.
Published: (2024)
by: Chen, Ying, et al.
Published: (2024)
Lightweight Model for Poultry Disease Detection from Fecal Images Using Multi-Color Space Feature Optimization and Machine Learning
by: Islam, A. K. M. Shoriful, et al.
Published: (2025)
by: Islam, A. K. M. Shoriful, et al.
Published: (2025)
Multi-modal Co-learning for Earth Observation: Enhancing single-modality models via modality collaboration
by: Mena, Francisco, et al.
Published: (2025)
by: Mena, Francisco, et al.
Published: (2025)
VideoCAD: A Dataset and Model for Learning Long-Horizon 3D CAD UI Interactions from Video
by: Man, Brandon, et al.
Published: (2025)
by: Man, Brandon, et al.
Published: (2025)
Skin Lesion Phenotyping via Nested Multi-modal Contrastive Learning
by: Christopoulos, Dionysis, et al.
Published: (2025)
by: Christopoulos, Dionysis, et al.
Published: (2025)
Training-Only Heterogeneous Image-Patch-Text Graph Supervision for Advancing Few-Shot Learning Adapters
by: Mohammad, Mohammed Rahman Sherif Khan, et al.
Published: (2026)
by: Mohammad, Mohammed Rahman Sherif Khan, et al.
Published: (2026)
Knowledge Graph Enhanced Generative Multi-modal Models for Class-Incremental Learning
by: Cao, Xusheng, et al.
Published: (2025)
by: Cao, Xusheng, et al.
Published: (2025)
DreamCAD: Scaling Multi-modal CAD Generation using Differentiable Parametric Surfaces
by: Khan, Mohammad Sadil, et al.
Published: (2026)
by: Khan, Mohammad Sadil, et al.
Published: (2026)
Exploring Machine Learning Engineering for Object Detection and Tracking by Unmanned Aerial Vehicle (UAV)
by: Guna, Aneesha, et al.
Published: (2024)
by: Guna, Aneesha, et al.
Published: (2024)
Classification of Multi-Parametric Body MRI Series Using Deep Learning
by: Kim, Boah, et al.
Published: (2025)
by: Kim, Boah, et al.
Published: (2025)
Cross-modal Affinity-aligned Multimodal Learning Analytics for Predicting Student Collaboration Satisfaction in Game-Based Learning
by: Tsai, Wen-Hsin, et al.
Published: (2026)
by: Tsai, Wen-Hsin, et al.
Published: (2026)
Advancing Stroke Risk Prediction Using a Multi-modal Foundation Model
by: Delgrange, Camille, et al.
Published: (2024)
by: Delgrange, Camille, et al.
Published: (2024)
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
by: Zhu, Mengdan, et al.
Published: (2025)
by: Zhu, Mengdan, et al.
Published: (2025)
Multi-modal Masked Siamese Network Improves Chest X-Ray Representation Learning
by: Shurrab, Saeed, et al.
Published: (2024)
by: Shurrab, Saeed, et al.
Published: (2024)
Proactive Agents for Multi-Turn Text-to-Image Generation Under Uncertainty
by: Hahn, Meera, et al.
Published: (2024)
by: Hahn, Meera, et al.
Published: (2024)
Concept Weaver: Enabling Multi-Concept Fusion in Text-to-Image Models
by: Kwon, Gihyun, et al.
Published: (2024)
by: Kwon, Gihyun, et al.
Published: (2024)
Propensity-to-Pay: Machine Learning for Estimating Prediction Uncertainty
by: Bashar, Md Abul, et al.
Published: (2020)
by: Bashar, Md Abul, et al.
Published: (2020)
Visual Hallucinations of Multi-modal Large Language Models
by: Huang, Wen, et al.
Published: (2024)
by: Huang, Wen, et al.
Published: (2024)
Rationale-Enhanced Decoding for Multi-modal Chain-of-Thought
by: Yamaguchi, Shin'ya, et al.
Published: (2025)
by: Yamaguchi, Shin'ya, et al.
Published: (2025)
EdgeFusion: On-Device Text-to-Image Generation
by: Castells, Thibault, et al.
Published: (2024)
by: Castells, Thibault, et al.
Published: (2024)
CellCLIP -- Learning Perturbation Effects in Cell Painting via Text-Guided Contrastive Learning
by: Lu, Mingyu, et al.
Published: (2025)
by: Lu, Mingyu, et al.
Published: (2025)
MulCPred: Learning Multi-modal Concepts for Explainable Pedestrian Action Prediction
by: Feng, Yan, et al.
Published: (2024)
by: Feng, Yan, et al.
Published: (2024)
GenSim2: Scaling Robot Data Generation with Multi-modal and Reasoning LLMs
by: Hua, Pu, et al.
Published: (2024)
by: Hua, Pu, et al.
Published: (2024)
Security Risk of Misalignment between Text and Image in Multi-modal Model
by: Wang, Xiaosen, et al.
Published: (2025)
by: Wang, Xiaosen, et al.
Published: (2025)
Contrastive Learning-based Multi Modal Architecture for Emoticon Prediction by Employing Image-Text Pairs
by: Pandey, Ananya, et al.
Published: (2024)
by: Pandey, Ananya, et al.
Published: (2024)
Generative Multi-modal Models are Good Class-Incremental Learners
by: Cao, Xusheng, et al.
Published: (2024)
by: Cao, Xusheng, et al.
Published: (2024)
UNIAA: A Unified Multi-modal Image Aesthetic Assessment Baseline and Benchmark
by: Zhou, Zhaokun, et al.
Published: (2024)
by: Zhou, Zhaokun, et al.
Published: (2024)
Enhancing Multi-task Learning Capability of Medical Generalist Foundation Model via Image-centric Multi-annotation Data
by: Zhu, Xun, et al.
Published: (2025)
by: Zhu, Xun, et al.
Published: (2025)
PATIMT-Bench: A Multi-Scenario Benchmark for Position-Aware Text Image Machine Translation in Large Vision-Language Models
by: Zhuang, Wanru, et al.
Published: (2025)
by: Zhuang, Wanru, et al.
Published: (2025)
Learning Hyperspectral Images with Curated Text Prompts for Efficient Multimodal Alignment
by: Chatterjee, Abhiroop, et al.
Published: (2025)
by: Chatterjee, Abhiroop, et al.
Published: (2025)
Hidden Bias in the Machine: Stereotypes in Text-to-Image Models
by: Porikli, Sedat, et al.
Published: (2025)
by: Porikli, Sedat, et al.
Published: (2025)
Attention-based Multi-modal Deep Learning Model of Spatio-temporal Crop Yield Prediction with Satellite, Soil and Climate Data
by: Shyam, Gopal Krishna, et al.
Published: (2026)
by: Shyam, Gopal Krishna, et al.
Published: (2026)
Towards Equitable ASD Diagnostics: A Comparative Study of Machine and Deep Learning Models Using Behavioral and Facial Data
by: Aledhari, Mohammed, et al.
Published: (2024)
by: Aledhari, Mohammed, et al.
Published: (2024)
Visual Explanations of Image-Text Representations via Multi-Modal Information Bottleneck Attribution
by: Wang, Ying, et al.
Published: (2023)
by: Wang, Ying, et al.
Published: (2023)
Tiny Machine Learning: Progress and Futures
by: Lin, Ji, et al.
Published: (2024)
by: Lin, Ji, et al.
Published: (2024)
Similar Items
-
BIKED++: A Multimodal Dataset of 1.4 Million Bicycle Image and Parametric CAD Designs
by: Regenwetter, Lyle, et al.
Published: (2024) -
Simultaneous Long-tailed Recognition and Multi-modal Fusion for Highly Imbalanced Multi-modal Data
by: Yoon, Heegeon, et al.
Published: (2026) -
Multi-modal Vision Pre-training for Medical Image Analysis
by: Rui, Shaohao, et al.
Published: (2024) -
Integrating Text and Image Pre-training for Multi-modal Algorithmic Reasoning
by: Zhang, Zijian, et al.
Published: (2024) -
Robust Disaster Assessment from Aerial Imagery Using Text-to-Image Synthetic Data
by: Kalluri, Tarun, et al.
Published: (2024)