Investigating the Impact of Model Width and Density on Generalization in Presence of Label Noise
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xue, Yihao, Whitecross, Kyle, Mirzasoleiman, Baharan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Challenges and Opportunities in Improving Worst-Group Generalization in Presence of Spurious Features
von: Joshi, Siddharth, et al.
Veröffentlicht: (2023)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2023)
Representations Shape Weak-to-Strong Generalization: Theoretical Insights and Empirical Predictions
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
Understanding the Robustness of Multi-modal Contrastive Learning to Distribution Shift
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
Investigating the Benefits of Projection Head for Representation Learning
von: Xue, Yihao, et al.
Veröffentlicht: (2024)
von: Xue, Yihao, et al.
Veröffentlicht: (2024)
Tuning the Implicit Regularizer of Masked Diffusion Language Models: Enhancing Generalization via Insights from $k$-Parity
von: Huang, Jianhao, et al.
Veröffentlicht: (2026)
von: Huang, Jianhao, et al.
Veröffentlicht: (2026)
Few-shot Adaptation to Distribution Shifts By Mixing Source and Target Embeddings
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
Graph Contrastive Learning under Heterophily via Graph Filters
von: Yang, Wenhan, et al.
Veröffentlicht: (2023)
von: Yang, Wenhan, et al.
Veröffentlicht: (2023)
Beyond What Seems Necessary: Hidden Gains from Scaling Training-Time Reasoning Length under Outcome Supervision
von: Xue, Yihao, et al.
Veröffentlicht: (2026)
von: Xue, Yihao, et al.
Veröffentlicht: (2026)
Data-Efficient Contrastive Self-supervised Learning: Most Beneficial Examples for Supervised Learning Contribute the Least
von: Joshi, Siddharth, et al.
Veröffentlicht: (2023)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2023)
LoRA is All You Need for Safety Alignment of Reasoning LLMs
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
Theoretical Perspectives on Data Quality and Synergistic Effects in Pre- and Post-Training Reasoning Models
von: Javanmard, Adel, et al.
Veröffentlicht: (2026)
von: Javanmard, Adel, et al.
Veröffentlicht: (2026)
Dataset Distillation via Knowledge Distillation: Towards Efficient Self-Supervised Pre-Training of Deep Networks
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
Understanding the Role of Training Data in Test-Time Scaling
von: Javanmard, Adel, et al.
Veröffentlicht: (2025)
von: Javanmard, Adel, et al.
Veröffentlicht: (2025)
Beyond Semantic Entropy: Boosting LLM Uncertainty Quantification with Pairwise Semantic Similarity
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
RecaLLM: Addressing the Lost-in-Thought Phenomenon with Explicit In-Context Retrieval
von: Whitecross, Kyle, et al.
Veröffentlicht: (2026)
von: Whitecross, Kyle, et al.
Veröffentlicht: (2026)
Do We Need All the Synthetic Data? Targeted Image Augmentation via Diffusion Models
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
Better Safe than Sorry: Pre-training CLIP against Targeted Data Poisoning and Backdoor Attacks
von: Yang, Wenhan, et al.
Veröffentlicht: (2023)
von: Yang, Wenhan, et al.
Veröffentlicht: (2023)
Changing the Training Data Distribution to Reduce Simplicity Bias Improves In-distribution Generalization
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
Data Distribution as a Lever for Guiding Optimizers Toward Superior Generalization in LLMs
von: Gangavarapu, Tushaar, et al.
Veröffentlicht: (2026)
von: Gangavarapu, Tushaar, et al.
Veröffentlicht: (2026)
SmallToLarge (S2L): Scalable Data Selection for Fine-tuning Large Language Models by Summarizing Training Trajectories of Small Models
von: Yang, Yu, et al.
Veröffentlicht: (2024)
von: Yang, Yu, et al.
Veröffentlicht: (2024)
Synthetic Text Generation for Training Large Language Models via Gradient Matching
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Dang, et al.
Veröffentlicht: (2025)
Data-Efficient Contrastive Language-Image Pretraining: Prioritizing Data Quality over Quantity
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2024)
Mini-batch Coresets for Memory-efficient Language Model Training on Data Mixtures
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
Identifying Spurious Biases Early in Training through the Lens of Simplicity Bias
von: Yang, Yu, et al.
Veröffentlicht: (2023)
von: Yang, Yu, et al.
Veröffentlicht: (2023)
Data Selection for Fine-tuning Vision Language Models via Cross Modal Alignment Trajectories
von: Naharas, Nilay, et al.
Veröffentlicht: (2025)
von: Naharas, Nilay, et al.
Veröffentlicht: (2025)
How Transformers Learn to Plan via Multi-Token Prediction
von: Huang, Jianhao, et al.
Veröffentlicht: (2026)
von: Huang, Jianhao, et al.
Veröffentlicht: (2026)
Scalable Out-of-distribution Robustness in the Presence of Unobserved Confounders
von: Prashant, Parjanya, et al.
Veröffentlicht: (2024)
von: Prashant, Parjanya, et al.
Veröffentlicht: (2024)
Verify when Uncertain: Beyond Self-Consistency in Black Box Hallucination Detection
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
von: Xue, Yihao, et al.
Veröffentlicht: (2025)
Balancing Label Imbalance in Federated Environments Using Only Mixup and Artificially-Labeled Noise
von: Sang, Kyle, et al.
Veröffentlicht: (2024)
von: Sang, Kyle, et al.
Veröffentlicht: (2024)
Learning Confident Classifiers in the Presence of Label Noise
von: Hashmi, Asma Ahmed, et al.
Veröffentlicht: (2023)
von: Hashmi, Asma Ahmed, et al.
Veröffentlicht: (2023)
Robust-GBDT: GBDT with Nonconvex Loss for Tabular Classification in the Presence of Label Noise and Class Imbalance
von: Luo, Jiaqi, et al.
Veröffentlicht: (2023)
von: Luo, Jiaqi, et al.
Veröffentlicht: (2023)
MM-GEN: Enhancing Task Performance Through Targeted Multimodal Data Curation
von: Joshi, Siddharth, et al.
Veröffentlicht: (2025)
von: Joshi, Siddharth, et al.
Veröffentlicht: (2025)
Impact of Label Noise on Learning Complex Features
von: Vashisht, Rahul, et al.
Veröffentlicht: (2024)
von: Vashisht, Rahul, et al.
Veröffentlicht: (2024)
Label-Noise Robust Diffusion Models
von: Na, Byeonghu, et al.
Veröffentlicht: (2024)
von: Na, Byeonghu, et al.
Veröffentlicht: (2024)
Active Label Refinement for Robust Training of Imbalanced Medical Image Classification Tasks in the Presence of High Label Noise
von: Khanal, Bidur, et al.
Veröffentlicht: (2024)
von: Khanal, Bidur, et al.
Veröffentlicht: (2024)
Generating the Ground Truth: Synthetic Data for Soft Label and Label Noise Research
von: de Vries, Sjoerd, et al.
Veröffentlicht: (2023)
von: de Vries, Sjoerd, et al.
Veröffentlicht: (2023)
Reduction-based Pseudo-label Generation for Instance-dependent Partial Label Learning
von: Qiao, Congyu, et al.
Veröffentlicht: (2024)
von: Qiao, Congyu, et al.
Veröffentlicht: (2024)
Label Noise: Ignorance Is Bliss
von: Zhu, Yilun, et al.
Veröffentlicht: (2024)
von: Zhu, Yilun, et al.
Veröffentlicht: (2024)
NoiseBench: Benchmarking the Impact of Real Label Noise on Named Entity Recognition
von: Merdjanovska, Elena, et al.
Veröffentlicht: (2024)
von: Merdjanovska, Elena, et al.
Veröffentlicht: (2024)
Do We Really Need Permutations? Impact of Model Width on Linear Mode Connectivity
von: Ito, Akira, et al.
Veröffentlicht: (2025)
von: Ito, Akira, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Challenges and Opportunities in Improving Worst-Group Generalization in Presence of Spurious Features
von: Joshi, Siddharth, et al.
Veröffentlicht: (2023) -
Representations Shape Weak-to-Strong Generalization: Theoretical Insights and Empirical Predictions
von: Xue, Yihao, et al.
Veröffentlicht: (2025) -
Understanding the Robustness of Multi-modal Contrastive Learning to Distribution Shift
von: Xue, Yihao, et al.
Veröffentlicht: (2023) -
Investigating the Benefits of Projection Head for Representation Learning
von: Xue, Yihao, et al.
Veröffentlicht: (2024) -
Tuning the Implicit Regularizer of Masked Diffusion Language Models: Enhancing Generalization via Insights from $k$-Parity
von: Huang, Jianhao, et al.
Veröffentlicht: (2026)