Changing the Training Data Distribution to Reduce Simplicity Bias Improves In-distribution Generalization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Nguyen, Dang, Haddad, Paymon, Gan, Eric, Mirzasoleiman, Baharan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Identifying Spurious Biases Early in Training through the Lens of Simplicity Bias
von: Yang, Yu, et al.
Veröffentlicht: (2023)
von: Yang, Yu, et al.
Veröffentlicht: (2023)
Better Safe than Sorry: Pre-training CLIP against Targeted Data Poisoning and Backdoor Attacks
von: Yang, Wenhan, et al.
Veröffentlicht: (2023)
von: Yang, Wenhan, et al.
Veröffentlicht: (2023)
Do We Need All the Synthetic Data? Targeted Image Augmentation via Diffusion Models
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
Data Selection for Fine-tuning Vision Language Models via Cross Modal Alignment Trajectories
von: Naharas, Nilay, et al.
Veröffentlicht: (2025)
von: Naharas, Nilay, et al.
Veröffentlicht: (2025)
Investigating the Benefits of Projection Head for Representation Learning
von: Xue, Yihao, et al.
Veröffentlicht: (2024)
von: Xue, Yihao, et al.
Veröffentlicht: (2024)
Data-Efficient Contrastive Language-Image Pretraining: Prioritizing Data Quality over Quantity
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
Understanding the Role of Training Data in Test-Time Scaling
von: Javanmard, Adel, et al.
Veröffentlicht: (2025)
von: Javanmard, Adel, et al.
Veröffentlicht: (2025)
Mini-batch Coresets for Memory-efficient Language Model Training on Data Mixtures
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
Data-Efficient Contrastive Self-supervised Learning: Most Beneficial Examples for Supervised Learning Contribute the Least
von: Joshi, Siddharth, et al.
Veröffentlicht: (2023)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2023)
Few-shot Adaptation to Distribution Shifts By Mixing Source and Target Embeddings
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
Theoretical Perspectives on Data Quality and Synergistic Effects in Pre- and Post-Training Reasoning Models
von: Javanmard, Adel, et al.
Veröffentlicht: (2026)
von: Javanmard, Adel, et al.
Veröffentlicht: (2026)
Tuning the Implicit Regularizer of Masked Diffusion Language Models: Enhancing Generalization via Insights from $k$-Parity
von: Huang, Jianhao, et al.
Veröffentlicht: (2026)
von: Huang, Jianhao, et al.
Veröffentlicht: (2026)
Data Distribution as a Lever for Guiding Optimizers Toward Superior Generalization in LLMs
von: Gangavarapu, Tushaar, et al.
Veröffentlicht: (2026)
von: Gangavarapu, Tushaar, et al.
Veröffentlicht: (2026)
Representations Shape Weak-to-Strong Generalization: Theoretical Insights and Empirical Predictions
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
Synthetic Simplicity: Unveiling Bias in Medical Data Augmentation
von: Babu, Krishan Agyakari Raja, et al.
Veröffentlicht: (2024)
von: Babu, Krishan Agyakari Raja, et al.
Veröffentlicht: (2024)
GReFEL: Geometry-Aware Reliable Facial Expression Learning under Bias and Imbalanced Data Distribution
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
von: Wasi, Azmine Toushik, et al.
Veröffentlicht: (2024)
VeCLIP: Improving CLIP Training via Visual-enriched Captions
von: Lai, Zhengfeng, et al.
Veröffentlicht: (2023)
von: Lai, Zhengfeng, et al.
Veröffentlicht: (2023)
When Less Is More: Simplicity Beats Complexity for Physics-Constrained InSAR Phase Unwrapping
von: Singh, Prabhjot, et al.
Veröffentlicht: (2026)
von: Singh, Prabhjot, et al.
Veröffentlicht: (2026)
Oscillation-Reduced MXFP4 Training for Vision Transformers
von: Chen, Yuxiang, et al.
Veröffentlicht: (2025)
von: Chen, Yuxiang, et al.
Veröffentlicht: (2025)
Robustmix: Improving Robustness by Regularizing the Frequency Bias of Deep Nets
von: Ngnawe, Jonas, et al.
Veröffentlicht: (2023)
von: Ngnawe, Jonas, et al.
Veröffentlicht: (2023)
Improving the Training of Rectified Flows
von: Lee, Sangyun, et al.
Veröffentlicht: (2024)
von: Lee, Sangyun, et al.
Veröffentlicht: (2024)
Data-Driven Analysis of Intersectional Bias in Image Classification: A Framework with Bias-Weighted Augmentation
von: Yesmin, Farjana
Veröffentlicht: (2025)
von: Yesmin, Farjana
Veröffentlicht: (2025)
Generate Any Scene: Scene Graph Driven Data Synthesis for Visual Generation Training
von: Gao, Ziqi, et al.
Veröffentlicht: (2024)
von: Gao, Ziqi, et al.
Veröffentlicht: (2024)
Toward an Artificial General Teacher: Procedural Geometry Data Generation and Visual Grounding with Vision-Language Models
von: Nguyen-Truong, Hai, et al.
Veröffentlicht: (2026)
von: Nguyen-Truong, Hai, et al.
Veröffentlicht: (2026)
Understanding, Accelerating, and Improving MeanFlow Training
von: Kim, Jin-Young, et al.
Veröffentlicht: (2025)
von: Kim, Jin-Young, et al.
Veröffentlicht: (2025)
ReMix: Training Generalized Person Re-identification on a Mixture of Data
von: Mamedov, Timur, et al.
Veröffentlicht: (2024)
von: Mamedov, Timur, et al.
Veröffentlicht: (2024)
Mitigating Individual Skin Tone Bias in Skin Lesion Classification through Distribution-Aware Reweighting
von: Paxton, Kuniko, et al.
Veröffentlicht: (2025)
von: Paxton, Kuniko, et al.
Veröffentlicht: (2025)
Memory-efficient Continual Learning with Neural Collapse Contrastive
von: Dang, Trung-Anh, et al.
Veröffentlicht: (2024)
von: Dang, Trung-Anh, et al.
Veröffentlicht: (2024)
Hierarchical Simplicity Bias of Neural Networks
von: Du, Zhehang
Veröffentlicht: (2023)
von: Du, Zhehang
Veröffentlicht: (2023)
ScoreMix: Synthetic Data Generation by Score Composition in Diffusion Models Improves Recognition
von: Rahimi, Parsa, et al.
Veröffentlicht: (2025)
von: Rahimi, Parsa, et al.
Veröffentlicht: (2025)
Adaptive Prediction Ensemble: Improving Out-of-Distribution Generalization of Motion Forecasting
von: Li, Jinning, et al.
Veröffentlicht: (2024)
von: Li, Jinning, et al.
Veröffentlicht: (2024)
Towards Robust Out-of-Distribution Generalization: Data Augmentation and Neural Architecture Search Approaches
von: Bai, Haoyue
Veröffentlicht: (2024)
von: Bai, Haoyue
Veröffentlicht: (2024)
Universal Multi-Domain Translation via Diffusion Routers
von: Kieu, Duc, et al.
Veröffentlicht: (2025)
von: Kieu, Duc, et al.
Veröffentlicht: (2025)
Distribution Shifts at Scale: Out-of-distribution Detection in Earth Observation
von: Ekim, Burak, et al.
Veröffentlicht: (2024)
von: Ekim, Burak, et al.
Veröffentlicht: (2024)
MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation
von: Joshi, Siddharth, et al.
Veröffentlicht: (2025)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2025)
Diminishing Stereotype Bias in Image Generation Model using Reinforcemenlent Learning Feedback
von: Chen, Xin, et al.
Veröffentlicht: (2024)
von: Chen, Xin, et al.
Veröffentlicht: (2024)
Improving Multi-Label Contrastive Learning by Leveraging Label Distribution
von: Chen, Ning, et al.
Veröffentlicht: (2025)
von: Chen, Ning, et al.
Veröffentlicht: (2025)
Identity Curvature Laplace Approximation for Improved Out-of-Distribution Detection
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2023)
von: Zhdanov, Maksim, et al.
Veröffentlicht: (2023)
Configuring Data Augmentations to Reduce Variance Shift in Positional Embedding of Vision Transformers
von: Kim, Bum Jun, et al.
Veröffentlicht: (2024)
von: Kim, Bum Jun, et al.
Veröffentlicht: (2024)
Classifier-to-Bias: Toward Unsupervised Automatic Bias Detection for Visual Classifiers
von: Guimard, Quentin, et al.
Veröffentlicht: (2025)
von: Guimard, Quentin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Identifying Spurious Biases Early in Training through the Lens of Simplicity Bias
von: Yang, Yu, et al.
Veröffentlicht: (2023) -
Better Safe than Sorry: Pre-training CLIP against Targeted Data Poisoning and Backdoor Attacks
von: Yang, Wenhan, et al.
Veröffentlicht: (2023) -
Do We Need All the Synthetic Data? Targeted Image Augmentation via Diffusion Models
von: Nguyen, Dang, et al.
Veröffentlicht: (2025) -
Data Selection for Fine-tuning Vision Language Models via Cross Modal Alignment Trajectories
von: Naharas, Nilay, et al.
Veröffentlicht: (2025) -
Investigating the Benefits of Projection Head for Representation Learning
von: Xue, Yihao, et al.
Veröffentlicht: (2024)