Diversify, Don't Fine-Tune: Scaling Up Visual Recognition Training with Synthetic Images
Fuente:
arXiv
Saved in:
| Main Authors: | Yu, Zhuoran, Zhu, Chenchen, Culatana, Sean, Krishnamoorthi, Raghuraman, Xiao, Fanyi, Lee, Yong Jae |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
How Multimodal LLMs Solve Image Tasks: A Lens on Visual Grounding, Task Reasoning, and Answer Decoding
by: Yu, Zhuoran, et al.
Published: (2025)
by: Yu, Zhuoran, et al.
Published: (2025)
Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time
by: Cheng, Jintao, et al.
Published: (2025)
by: Cheng, Jintao, et al.
Published: (2025)
Don't Just Fine-tune the Agent, Tune the Environment
by: Lu, Siyuan, et al.
Published: (2025)
by: Lu, Siyuan, et al.
Published: (2025)
Large Pre-Training Datasets Don't Always Guarantee Robustness after Fine-Tuning
by: Hwang, Jaedong, et al.
Published: (2024)
by: Hwang, Jaedong, et al.
Published: (2024)
EdgeTAM: On-Device Track Anything Model
by: Zhou, Chong, et al.
Published: (2025)
by: Zhou, Chong, et al.
Published: (2025)
Language Models Don't Learn the Physical Manifestation of Language
by: Lee, Bruce W., et al.
Published: (2024)
by: Lee, Bruce W., et al.
Published: (2024)
Don't Forget the Nonlinearity: Unlocking Activation Functions in Efficient Fine-Tuning
by: Yin, Bo, et al.
Published: (2025)
by: Yin, Bo, et al.
Published: (2025)
Surely Large Multimodal Models (Don't) Excel in Visual Species Recognition?
by: Liu, Tian, et al.
Published: (2025)
by: Liu, Tian, et al.
Published: (2025)
Don't Retrieve, Generate: Prompting LLMs for Synthetic Training Data in Dense Retrieval
by: Sinha, Aarush
Published: (2025)
by: Sinha, Aarush
Published: (2025)
You Don't Need All Attentions: Distributed Dynamic Fine-Tuning for Foundation Models
by: Ding, Shiwei, et al.
Published: (2025)
by: Ding, Shiwei, et al.
Published: (2025)
Don't Let Your Likert Scales Grow Up To Be Visual Analog Scales: Understanding the Relationship Between Number of Response Categories and Measurement Error
by: Sun, Siqi, et al.
Published: (2025)
by: Sun, Siqi, et al.
Published: (2025)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
by: Cha, Sungguk, et al.
Published: (2024)
by: Cha, Sungguk, et al.
Published: (2024)
Fine-Tuned LLMs Know They Don't Know: A Parameter-Efficient Approach to Recovering Honesty
by: Shi, Zeyu, et al.
Published: (2025)
by: Shi, Zeyu, et al.
Published: (2025)
Adapt, But Don't Forget: Fine-Tuning and Contrastive Routing for Lane Detection under Distribution Shift
by: Khan, Mohammed Abdul Hafeez, et al.
Published: (2025)
by: Khan, Mohammed Abdul Hafeez, et al.
Published: (2025)
Don't Fine-Tune, Decode: Syntax Error-Free Tool Use via Constrained Decoding
by: Zhang, Kexun, et al.
Published: (2023)
by: Zhang, Kexun, et al.
Published: (2023)
Fine-Tune, Don't Prompt, Your Language Model to Identify Biased Language in Clinical Notes
by: Landi, Isotta, et al.
Published: (2026)
by: Landi, Isotta, et al.
Published: (2026)
"Don't Mess Up My Algorithm": Phatic Communication and Algorithmic Contagion in Meme Sharing
by: Song, Ji Eun, et al.
Published: (2026)
by: Song, Ji Eun, et al.
Published: (2026)
Communication Efficient Distributed Training with Distributed Lion
by: Liu, Bo, et al.
Published: (2024)
by: Liu, Bo, et al.
Published: (2024)
Grow, Don't Overwrite: Fine-tuning Without Forgetting
by: Adila, Dyah, et al.
Published: (2026)
by: Adila, Dyah, et al.
Published: (2026)
Vision Transformers Don't Need Trained Registers
by: Jiang, Nick, et al.
Published: (2025)
by: Jiang, Nick, et al.
Published: (2025)
What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"
by: Lee, Joosung, et al.
Published: (2026)
by: Lee, Joosung, et al.
Published: (2026)
Mix, Don't Tune: Bilingual Pre-Training Outperforms Hyperparameter Search in Data-Constrained Settings
by: Jeha, Paul, et al.
Published: (2026)
by: Jeha, Paul, et al.
Published: (2026)
R-Tuning: Instructing Large Language Models to Say `I Don't Know'
by: Zhang, Hanning, et al.
Published: (2023)
by: Zhang, Hanning, et al.
Published: (2023)
SqueezeSAM: User friendly mobile interactive segmentation
by: Varadarajan, Balakrishnan, et al.
Published: (2023)
by: Varadarajan, Balakrishnan, et al.
Published: (2023)
Don't be salesmen
Published: (1997)
Published: (1997)
Honest AI: Fine-Tuning "Small" Language Models to Say "I Don't Know", and Reducing Hallucination in RAG
by: Chen, Xinxi, et al.
Published: (2024)
by: Chen, Xinxi, et al.
Published: (2024)
Synthetic Data Generation for Training Diversified Commonsense Reasoning Models
by: Zhang, Tianhui, et al.
Published: (2026)
by: Zhang, Tianhui, et al.
Published: (2026)
Improved Baselines with Visual Instruction Tuning
by: Liu, Haotian, et al.
Published: (2023)
by: Liu, Haotian, et al.
Published: (2023)
Don't Mesh with Me: Generating Constructive Solid Geometry Instead of Meshes by Fine-Tuning a Code-Generation LLM
by: Mews, Maximilian, et al.
Published: (2024)
by: Mews, Maximilian, et al.
Published: (2024)
Don’t Burn it Here
by: Walsh, Ed, et al.
Published: (2026)
by: Walsh, Ed, et al.
Published: (2026)
Don't Look Away
by: Cohen, Brianne
Published: (2023)
by: Cohen, Brianne
Published: (2023)
Don't Forget Imagination!
by: Vityaev, Evgenii E., et al.
Published: (2025)
by: Vityaev, Evgenii E., et al.
Published: (2025)
Similar Items
-
How Multimodal LLMs Solve Image Tasks: A Lens on Visual Grounding, Task Reasoning, and Answer Decoding
by: Yu, Zhuoran, et al.
Published: (2025) -
Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time
by: Cheng, Jintao, et al.
Published: (2025) -
Don't Just Fine-tune the Agent, Tune the Environment
by: Lu, Siyuan, et al.
Published: (2025) -
Large Pre-Training Datasets Don't Always Guarantee Robustness after Fine-Tuning
by: Hwang, Jaedong, et al.
Published: (2024) -
EdgeTAM: On-Device Track Anything Model
by: Zhou, Chong, et al.
Published: (2025)