Investigating the Impact of Large-Scale Pre-training on Nutritional Content Estimation from 2D Images
Fuente:
arXiv
Saved in:
| Main Authors: | Andrade, Michele, Silva, Guilherme A. L., Santos, Valéria, Moreira, Gladston, Luz, Eduardo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Systematic Review of ECG Arrhythmia Classification: Adherence to Standards, Fair Evaluation, and Embedded Feasibility
by: Silva, Guilherme, et al.
Published: (2025)
by: Silva, Guilherme, et al.
Published: (2025)
PD-Loss: Proxy-Decidability for Efficient Metric Learning
by: Silva, Pedro, et al.
Published: (2025)
by: Silva, Pedro, et al.
Published: (2025)
Large-Scale 3D Medical Image Pre-training with Geometric Context Priors
by: Wu, Linshan, et al.
Published: (2024)
by: Wu, Linshan, et al.
Published: (2024)
Leveraging Visibility Graphs for Enhanced Arrhythmia Classification with Graph Convolutional Networks
by: Oliveira, Rafael F., et al.
Published: (2024)
by: Oliveira, Rafael F., et al.
Published: (2024)
Scalable Pre-training of Large Autoregressive Image Models
by: El-Nouby, Alaaeldin, et al.
Published: (2024)
by: El-Nouby, Alaaeldin, et al.
Published: (2024)
Pre-Training for 3D Hand Pose Estimation with Contrastive Learning on Large-Scale Hand Images in the Wild
by: Lin, Nie, et al.
Published: (2024)
by: Lin, Nie, et al.
Published: (2024)
Maximum Dispersion, Maximum Concentration: Enhancing the Quality of MOP Solutions
by: Moreira, Gladston, et al.
Published: (2025)
by: Moreira, Gladston, et al.
Published: (2025)
SiMHand: Mining Similar Hands for Large-Scale 3D Hand Pose Pre-training
by: Lin, Nie, et al.
Published: (2025)
by: Lin, Nie, et al.
Published: (2025)
Towards Scalable Language-Image Pre-training for 3D Medical Imaging
by: Zhao, Chenhui, et al.
Published: (2025)
by: Zhao, Chenhui, et al.
Published: (2025)
Scaling Backwards: Minimal Synthetic Pre-training?
by: Nakamura, Ryo, et al.
Published: (2024)
by: Nakamura, Ryo, et al.
Published: (2024)
NimbleD: Enhancing Self-supervised Monocular Depth Estimation with Pseudo-labels and Large-scale Video Pre-training
by: Luginov, Albert, et al.
Published: (2024)
by: Luginov, Albert, et al.
Published: (2024)
Advancing Multinational License Plate Recognition Through Synthetic and Real Data Fusion: A Comprehensive Evaluation
by: Laroca, Rayson, et al.
Published: (2026)
by: Laroca, Rayson, et al.
Published: (2026)
Primitive Geometry Segment Pre-training for 3D Medical Image Segmentation
by: Tadokoro, Ryu, et al.
Published: (2024)
by: Tadokoro, Ryu, et al.
Published: (2024)
Should VLMs be Pre-trained with Image Data?
by: Keh, Sedrick, et al.
Published: (2025)
by: Keh, Sedrick, et al.
Published: (2025)
Evaluating Pre-trained Convolutional Neural Networks and Foundation Models as Feature Extractors for Content-based Medical Image Retrieval
by: Mahbod, Amirreza, et al.
Published: (2024)
by: Mahbod, Amirreza, et al.
Published: (2024)
Sculpting Holistic 3D Representation in Contrastive Language-Image-3D Pre-training
by: Gao, Yipeng, et al.
Published: (2023)
by: Gao, Yipeng, et al.
Published: (2023)
Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models
by: Lai, Zhengfeng, et al.
Published: (2024)
by: Lai, Zhengfeng, et al.
Published: (2024)
Muskie: Multi-view Masked Image Modeling for 3D Vision Pre-training
by: Li, Wenyu, et al.
Published: (2025)
by: Li, Wenyu, et al.
Published: (2025)
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training
by: Cao, Anjia, et al.
Published: (2024)
by: Cao, Anjia, et al.
Published: (2024)
Large-scale and Fine-grained Vision-language Pre-training for Enhanced CT Image Understanding
by: Shui, Zhongyi, et al.
Published: (2025)
by: Shui, Zhongyi, et al.
Published: (2025)
Multimodal Autoregressive Pre-training of Large Vision Encoders
by: Fini, Enrico, et al.
Published: (2024)
by: Fini, Enrico, et al.
Published: (2024)
Pre-training of Lightweight Vision Transformers on Small Datasets with Minimally Scaled Images
by: Tan, Jen Hong
Published: (2024)
by: Tan, Jen Hong
Published: (2024)
Pre-training Everywhere: Parameter-Efficient Fine-Tuning for Medical Image Analysis via Target Parameter Pre-training
by: Lei, Xingliang, et al.
Published: (2024)
by: Lei, Xingliang, et al.
Published: (2024)
MaskHOI: Robust 3D Hand-Object Interaction Estimation via Masked Pre-training
by: Xie, Yuechen, et al.
Published: (2025)
by: Xie, Yuechen, et al.
Published: (2025)
P3P: Pseudo-3D Pre-training for Scaling 3D Voxel-based Masked Autoencoders
by: Chen, Xuechao, et al.
Published: (2024)
by: Chen, Xuechao, et al.
Published: (2024)
4D Visual Pre-training for Robot Learning
by: Hou, Chengkai, et al.
Published: (2025)
by: Hou, Chengkai, et al.
Published: (2025)
Generic Knowledge Boosted Pre-training For Remote Sensing Images
by: Huang, Ziyue, et al.
Published: (2024)
by: Huang, Ziyue, et al.
Published: (2024)
PLIP: Language-Image Pre-training for Person Representation Learning
by: Zuo, Jialong, et al.
Published: (2023)
by: Zuo, Jialong, et al.
Published: (2023)
Universal Image Restoration Pre-training via Degradation Classification
by: Hu, JiaKui, et al.
Published: (2025)
by: Hu, JiaKui, et al.
Published: (2025)
DreamLIP: Language-Image Pre-training with Long Captions
by: Zheng, Kecheng, et al.
Published: (2024)
by: Zheng, Kecheng, et al.
Published: (2024)
Let ViT Speak: Generative Language-Image Pre-training
by: Fang, Yan, et al.
Published: (2026)
by: Fang, Yan, et al.
Published: (2026)
Advancing ALS Applications with Large-Scale Pre-training: Dataset Development and Downstream Assessment
by: Xiu, Haoyi, et al.
Published: (2025)
by: Xiu, Haoyi, et al.
Published: (2025)
Low-hallucination Synthetic Captions for Large-Scale Vision-Language Model Pre-training
by: Zhang, Xinsong, et al.
Published: (2025)
by: Zhang, Xinsong, et al.
Published: (2025)
Large-scale Pre-training for Grounded Video Caption Generation
by: Kazakos, Evangelos, et al.
Published: (2025)
by: Kazakos, Evangelos, et al.
Published: (2025)
MatchAnything: Universal Cross-Modality Image Matching with Large-Scale Pre-Training
by: He, Xingyi, et al.
Published: (2025)
by: He, Xingyi, et al.
Published: (2025)
KaoLRM: Repurposing Pre-trained Large Reconstruction Models for Parametric 3D Face Reconstruction
by: Zhu, Qingtian, et al.
Published: (2026)
by: Zhu, Qingtian, et al.
Published: (2026)
Scaling up Multimodal Pre-training for Sign Language Understanding
by: Zhou, Wengang, et al.
Published: (2024)
by: Zhou, Wengang, et al.
Published: (2024)
GPT4Image: Large Pre-trained Models Help Vision Models Learn Better on Perception Task
by: Ding, Ning, et al.
Published: (2023)
by: Ding, Ning, et al.
Published: (2023)
Scaling Pre-training to One Hundred Billion Data for Vision Language Models
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
Reinforcing Pre-trained Models Using Counterfactual Images
by: Li, Xiang, et al.
Published: (2024)
by: Li, Xiang, et al.
Published: (2024)
Similar Items
-
A Systematic Review of ECG Arrhythmia Classification: Adherence to Standards, Fair Evaluation, and Embedded Feasibility
by: Silva, Guilherme, et al.
Published: (2025) -
PD-Loss: Proxy-Decidability for Efficient Metric Learning
by: Silva, Pedro, et al.
Published: (2025) -
Large-Scale 3D Medical Image Pre-training with Geometric Context Priors
by: Wu, Linshan, et al.
Published: (2024) -
Leveraging Visibility Graphs for Enhanced Arrhythmia Classification with Graph Convolutional Networks
by: Oliveira, Rafael F., et al.
Published: (2024) -
Scalable Pre-training of Large Autoregressive Image Models
by: El-Nouby, Alaaeldin, et al.
Published: (2024)