Is Large-Scale Pretraining the Secret to Good Domain Generalization?
Fuente:
arXiv
Saved in:
| Main Authors: | Teterwak, Piotr, Saito, Kuniaki, Tsiligkaridis, Theodoros, Plummer, Bryan A., Saenko, Kate |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ERM++: An Improved Baseline for Domain Generalization
by: Teterwak, Piotr, et al.
Published: (2023)
by: Teterwak, Piotr, et al.
Published: (2023)
OP-LoRA: The Blessing of Dimensionality
by: Teterwak, Piotr, et al.
Published: (2024)
by: Teterwak, Piotr, et al.
Published: (2024)
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks
by: Qraitem, Maan, et al.
Published: (2024)
by: Qraitem, Maan, et al.
Published: (2024)
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
by: Qraitem, Maan, et al.
Published: (2023)
by: Qraitem, Maan, et al.
Published: (2023)
SLANT: Spurious Logo ANalysis Toolkit
by: Qraitem, Maan, et al.
Published: (2024)
by: Qraitem, Maan, et al.
Published: (2024)
Web Artifact Attacks Disrupt Vision Language Models
by: Qraitem, Maan, et al.
Published: (2025)
by: Qraitem, Maan, et al.
Published: (2025)
CLAMP: Contrastive LAnguage Model Prompt-tuning
by: Teterwak, Piotr, et al.
Published: (2023)
by: Teterwak, Piotr, et al.
Published: (2023)
Scaling Up Temporal Domain Generalization via Temporal Experts Averaging
by: Liu, Aoming, et al.
Published: (2025)
by: Liu, Aoming, et al.
Published: (2025)
Multimodal Unsupervised Domain Generalization by Retrieving Across the Modality Gap
by: Liao, Christopher, et al.
Published: (2024)
by: Liao, Christopher, et al.
Published: (2024)
Noise-Aware Generalization: Robustness to In-Domain Noise and Out-of-Domain Generalization
by: Wang, Siqi, et al.
Published: (2025)
by: Wang, Siqi, et al.
Published: (2025)
Tell Me What's Next: Textual Foresight for Generic UI Representations
by: Burns, Andrea, et al.
Published: (2024)
by: Burns, Andrea, et al.
Published: (2024)
Federated Adversarial Domain Adaptation
by: Peng, Xingchao, et al.
Published: (2019)
by: Peng, Xingchao, et al.
Published: (2019)
Image-Caption Encoding for Improving Zero-Shot Generalization
by: Yu, Eric Yang, et al.
Published: (2024)
by: Yu, Eric Yang, et al.
Published: (2024)
Concept Arithmetics for Circumventing Concept Inhibition in Diffusion Models
by: Petsiuk, Vitali, et al.
Published: (2024)
by: Petsiuk, Vitali, et al.
Published: (2024)
Weak-to-Strong Compositional Learning from Generative Models for Language-based Object Detection
by: Park, Kwanyong, et al.
Published: (2024)
by: Park, Kwanyong, et al.
Published: (2024)
RECAST: Reparameterized, Compact weight Adaptation for Sequential Tasks
by: Tasnim, Nazia, et al.
Published: (2024)
by: Tasnim, Nazia, et al.
Published: (2024)
Quantified Task Misalignment to Inform PEFT: An Exploration of Domain Generalization and Catastrophic Forgetting in CLIP
by: Niss, Laura, et al.
Published: (2024)
by: Niss, Laura, et al.
Published: (2024)
Pretrained Image-Text Models are Secretly Video Captioners
by: Zhang, Chunhui, et al.
Published: (2025)
by: Zhang, Chunhui, et al.
Published: (2025)
ChA-MAEViT: Unifying Channel-Aware Masked Autoencoders and Multi-Channel Vision Transformers for Improved Cross-Channel Learning
by: Pham, Chau, et al.
Published: (2025)
by: Pham, Chau, et al.
Published: (2025)
Stable Diffusion Models are Secretly Good at Visual In-Context Learning
by: Oorloff, Trevine, et al.
Published: (2025)
by: Oorloff, Trevine, et al.
Published: (2025)
Descriptor and Word Soups: Overcoming the Parameter Efficiency Accuracy Tradeoff for Out-of-Distribution Few-shot Learning
by: Liao, Christopher, et al.
Published: (2023)
by: Liao, Christopher, et al.
Published: (2023)
Scale-Wise VAR is Secretly Discrete Diffusion
by: Kumar, Amandeep, et al.
Published: (2025)
by: Kumar, Amandeep, et al.
Published: (2025)
Understanding Domain Generalization: A Noise Robustness Perspective
by: Qiao, Rui, et al.
Published: (2024)
by: Qiao, Rui, et al.
Published: (2024)
Multi-axis Analysis of Image Manipulation Localization
by: Nichols, Keanu, et al.
Published: (2026)
by: Nichols, Keanu, et al.
Published: (2026)
LargeAD: Large-Scale Cross-Sensor Data Pretraining for Autonomous Driving
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
The Inter-Intra Modal Measure: A Predictive Lens on Fine-Tuning Outcomes in Vision-Language Models
by: Niss, Laura, et al.
Published: (2024)
by: Niss, Laura, et al.
Published: (2024)
Koala: Key frame-conditioned long video-LLM
by: Tan, Reuben, et al.
Published: (2024)
by: Tan, Reuben, et al.
Published: (2024)
You Don't Need Domain-Specific Data Augmentations When Scaling Self-Supervised Learning
by: Moutakanni, Théo, et al.
Published: (2024)
by: Moutakanni, Théo, et al.
Published: (2024)
LoTUS: Large-Scale Machine Unlearning with a Taste of Uncertainty
by: Spartalis, Christoforos N., et al.
Published: (2025)
by: Spartalis, Christoforos N., et al.
Published: (2025)
Rethinking Generative Image Pretraining: How Far Are We From Scaling Up Next-Pixel Prediction?
by: Yan, Xinchen, et al.
Published: (2025)
by: Yan, Xinchen, et al.
Published: (2025)
KiVA: Kid-inspired Visual Analogies for Testing Large Multimodal Models
by: Yiu, Eunice, et al.
Published: (2024)
by: Yiu, Eunice, et al.
Published: (2024)
Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos
by: Luo, Hao, et al.
Published: (2025)
by: Luo, Hao, et al.
Published: (2025)
VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition
by: Yadav, Tanush, et al.
Published: (2026)
by: Yadav, Tanush, et al.
Published: (2026)
Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining
by: Xiong, Weimin, et al.
Published: (2026)
by: Xiong, Weimin, et al.
Published: (2026)
Why Fine-grained Labels in Pretraining Benefit Generalization?
by: Hong, Guan Zhe, et al.
Published: (2024)
by: Hong, Guan Zhe, et al.
Published: (2024)
What Secrets Do Your Manifolds Hold? Understanding the Local Geometry of Generative Models
by: Humayun, Ahmed Imtiaz, et al.
Published: (2024)
by: Humayun, Ahmed Imtiaz, et al.
Published: (2024)
Impact of Pretraining Word Co-occurrence on Compositional Generalization in Multimodal Models
by: Qu, Helen, et al.
Published: (2025)
by: Qu, Helen, et al.
Published: (2025)
Towards Safer Mobile Agents: Scalable Generation and Evaluation of Diverse Scenarios for VLMs
by: Taniguchi, Takara, et al.
Published: (2026)
by: Taniguchi, Takara, et al.
Published: (2026)
Box Pose and Shape Estimation and Domain Adaptation for Large-Scale Warehouse Automation
by: Yu, Xihang, et al.
Published: (2025)
by: Yu, Xihang, et al.
Published: (2025)
Pretrained Visual Uncertainties
by: Kirchhof, Michael, et al.
Published: (2024)
by: Kirchhof, Michael, et al.
Published: (2024)
Similar Items
-
ERM++: An Improved Baseline for Domain Generalization
by: Teterwak, Piotr, et al.
Published: (2023) -
OP-LoRA: The Blessing of Dimensionality
by: Teterwak, Piotr, et al.
Published: (2024) -
Vision-LLMs Can Fool Themselves with Self-Generated Typographic Attacks
by: Qraitem, Maan, et al.
Published: (2024) -
From Fake to Real: Pretraining on Balanced Synthetic Images to Prevent Spurious Correlations in Image Recognition
by: Qraitem, Maan, et al.
Published: (2023) -
SLANT: Spurious Logo ANalysis Toolkit
by: Qraitem, Maan, et al.
Published: (2024)