LLMs-based Augmentation for Domain Adaptation in Long-tailed Food Datasets
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Qing, Ngo, Chong-Wah, Lim, Ee-Peng, Sun, Qianru |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Unbiased Cross-Modal Representation Learning for Food Image-to-Recipe Retrieval
by: Wang, Qing, et al.
Published: (2025)
by: Wang, Qing, et al.
Published: (2025)
Mitigating Cross-modal Representation Bias for Multicultural Image-to-Recipe Retrieval
by: Wang, Qing, et al.
Published: (2025)
by: Wang, Qing, et al.
Published: (2025)
OVFoodSeg: Elevating Open-Vocabulary Food Image Segmentation via Image-Informed Textual Representation
by: Wu, Xiongwei, et al.
Published: (2024)
by: Wu, Xiongwei, et al.
Published: (2024)
Advancing Food Nutrition Estimation via Visual-Ingredient Feature Fusion
by: Qi, Huiyan, et al.
Published: (2025)
by: Qi, Huiyan, et al.
Published: (2025)
Interpretable Embedding for Ad-hoc Video Search
by: Wu, Jiaxin, et al.
Published: (2024)
by: Wu, Jiaxin, et al.
Published: (2024)
FoodLMM: A Versatile Food Assistant using Large Multi-modal Model
by: Yin, Yuehao, et al.
Published: (2023)
by: Yin, Yuehao, et al.
Published: (2023)
Retrieval Augmented Recipe Generation
by: Liu, Guoshan, et al.
Published: (2024)
by: Liu, Guoshan, et al.
Published: (2024)
Improving Interpretable Embeddings for Ad-hoc Video Search with Generative Captions and Multi-word Concept Bank
by: Wu, Jiaxin, et al.
Published: (2024)
by: Wu, Jiaxin, et al.
Published: (2024)
Seeing Culture: A Benchmark for Visual Reasoning and Grounding
by: Satar, Burak, et al.
Published: (2025)
by: Satar, Burak, et al.
Published: (2025)
PosMLP-Video: Spatial and Temporal Relative Position Encoding for Efficient Video Recognition
by: Hao, Yanbin, et al.
Published: (2024)
by: Hao, Yanbin, et al.
Published: (2024)
RoDE: Linear Rectified Mixture of Diverse Experts for Food Large Multi-Modal Models
by: Jiao, Pengkun, et al.
Published: (2024)
by: Jiao, Pengkun, et al.
Published: (2024)
Synthetic Data Augmentation using Pre-trained Diffusion Models for Long-tailed Food Image Classification
by: Koh, GaYeon, et al.
Published: (2025)
by: Koh, GaYeon, et al.
Published: (2025)
Diversified Augmentation with Domain Adaptation for Debiased Video Temporal Grounding
by: Ren, Junlong, et al.
Published: (2025)
by: Ren, Junlong, et al.
Published: (2025)
Subtyping Breast Lesions via Generative Augmentation based Long-tailed Recognition in Ultrasound
by: Chen, Shijing, et al.
Published: (2025)
by: Chen, Shijing, et al.
Published: (2025)
Long-tailed Species Recognition in the NACTI Wildlife Dataset
by: Liu, Zehua, et al.
Published: (2025)
by: Liu, Zehua, et al.
Published: (2025)
Semi-Supervised Domain Adaptation Using Target-Oriented Domain Augmentation for 3D Object Detection
by: Kim, Yecheol, et al.
Published: (2024)
by: Kim, Yecheol, et al.
Published: (2024)
Balanced Representation Learning for Long-tailed Skeleton-based Action Recognition
by: Liu, Hongda, et al.
Published: (2023)
by: Liu, Hongda, et al.
Published: (2023)
Mitigating Long-tail Distribution in Oracle Bone Inscriptions: Dataset, Model, and Benchmark
by: Li, Jinhao, et al.
Published: (2025)
by: Li, Jinhao, et al.
Published: (2025)
MORDA: A Synthetic Dataset to Facilitate Adaptation of Object Detectors to Unseen Real-target Domain While Preserving Performance on Real-source Domain
by: Lim, Hojun, et al.
Published: (2025)
by: Lim, Hojun, et al.
Published: (2025)
Towards Improved Proxy-based Deep Metric Learning via Data-Augmented Domain Adaptation
by: Ren, Li, et al.
Published: (2024)
by: Ren, Li, et al.
Published: (2024)
Semantic Data Augmentation for Long-tailed Facial Expression Recognition
by: Li, Zijian, et al.
Published: (2024)
by: Li, Zijian, et al.
Published: (2024)
Class Agnostic Instance-level Descriptor for Visual Instance Search
by: Sun, Qi-Ying, et al.
Published: (2025)
by: Sun, Qi-Ying, et al.
Published: (2025)
SciLT: Long-tailed Image Classification under Scientific Image Domains
by: Chen, Jiahao, et al.
Published: (2026)
by: Chen, Jiahao, et al.
Published: (2026)
Weakly-Supervised Semantic Segmentation with Image-Level Labels: from Traditional Models to Foundation Models
by: Chen, Zhaozheng, et al.
Published: (2023)
by: Chen, Zhaozheng, et al.
Published: (2023)
Video Anomaly Detection and Explanation via Large Language Models
by: Lv, Hui, et al.
Published: (2024)
by: Lv, Hui, et al.
Published: (2024)
RCA: Region Conditioned Adaptation for Visual Abductive Reasoning
by: Zhang, Hao, et al.
Published: (2023)
by: Zhang, Hao, et al.
Published: (2023)
LTGC: Long-tail Recognition via Leveraging LLMs-driven Generated Content
by: Zhao, Qihao, et al.
Published: (2024)
by: Zhao, Qihao, et al.
Published: (2024)
Frequency-based Matcher for Long-tailed Semantic Segmentation
by: Li, Shan, et al.
Published: (2024)
by: Li, Shan, et al.
Published: (2024)
Don't Deceive Me: Mitigating Gaslighting through Attention Reallocation in LMMs
by: Jiao, Pengkun, et al.
Published: (2025)
by: Jiao, Pengkun, et al.
Published: (2025)
From Holistic to Localized: Local Enhanced Adapters for Efficient Visual Instruction Fine-Tuning
by: Jiao, Pengkun, et al.
Published: (2024)
by: Jiao, Pengkun, et al.
Published: (2024)
Self-supervised Anomaly Detection Pretraining Enhances Long-tail ECG Diagnosis
by: Jiang, Aofan, et al.
Published: (2024)
by: Jiang, Aofan, et al.
Published: (2024)
CookingDiffusion: Cooking Procedural Image Generation with Stable Diffusion
by: Wang, Yuan, et al.
Published: (2025)
by: Wang, Yuan, et al.
Published: (2025)
HOT: Harmonic-Constrained Optimal Transport for Remote Photoplethysmography Domain Adaptation
by: Nguyen, Ba-Thinh, et al.
Published: (2026)
by: Nguyen, Ba-Thinh, et al.
Published: (2026)
Stratified Domain Adaptation: A Progressive Self-Training Approach for Scene Text Recognition
by: Le, Kha Nhat, et al.
Published: (2024)
by: Le, Kha Nhat, et al.
Published: (2024)
Adaptive Hardness-driven Augmentation and Alignment Strategies for Multi-Source Domain Adaptations
by: Yuxiang, Yang, et al.
Published: (2025)
by: Yuxiang, Yang, et al.
Published: (2025)
Rethinking Long-tailed Dataset Distillation: A Uni-Level Framework with Unbiased Recovery and Relabeling
by: Cui, Xiao, et al.
Published: (2025)
by: Cui, Xiao, et al.
Published: (2025)
DATR: Unsupervised Domain Adaptive Detection Transformer with Dataset-Level Adaptation and Prototypical Alignment
by: Han, Jianhong, et al.
Published: (2024)
by: Han, Jianhong, et al.
Published: (2024)
Hi3D: Pursuing High-Resolution Image-to-3D Generation with Video Diffusion Models
by: Yang, Haibo, et al.
Published: (2024)
by: Yang, Haibo, et al.
Published: (2024)
Semantic-guided Fine-tuning of Foundation Model for Long-tailed Visual Recognition
by: Peng, Yufei, et al.
Published: (2025)
by: Peng, Yufei, et al.
Published: (2025)
Label-Augmented Dataset Distillation
by: Kang, Seoungyoon, et al.
Published: (2024)
by: Kang, Seoungyoon, et al.
Published: (2024)
Similar Items
-
Towards Unbiased Cross-Modal Representation Learning for Food Image-to-Recipe Retrieval
by: Wang, Qing, et al.
Published: (2025) -
Mitigating Cross-modal Representation Bias for Multicultural Image-to-Recipe Retrieval
by: Wang, Qing, et al.
Published: (2025) -
OVFoodSeg: Elevating Open-Vocabulary Food Image Segmentation via Image-Informed Textual Representation
by: Wu, Xiongwei, et al.
Published: (2024) -
Advancing Food Nutrition Estimation via Visual-Ingredient Feature Fusion
by: Qi, Huiyan, et al.
Published: (2025) -
Interpretable Embedding for Ad-hoc Video Search
by: Wu, Jiaxin, et al.
Published: (2024)