Ask Your Distribution Shift if Pre-Training is Right for You
Fuente:
arXiv
Saved in:
| Main Authors: | Cohen-Wang, Benjamin, Vendrow, Joshua, Madry, Aleksander |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do Large Language Model Benchmarks Test Reliability?
by: Vendrow, Joshua, et al.
Published: (2025)
by: Vendrow, Joshua, et al.
Published: (2025)
Learning to Attribute with Attention
by: Cohen-Wang, Benjamin, et al.
Published: (2025)
by: Cohen-Wang, Benjamin, et al.
Published: (2025)
ContextCite: Attributing Model Generation to Context
by: Cohen-Wang, Benjamin, et al.
Published: (2024)
by: Cohen-Wang, Benjamin, et al.
Published: (2024)
DsDm: Model-Aware Dataset Selection with Datamodels
by: Engstrom, Logan, et al.
Published: (2024)
by: Engstrom, Logan, et al.
Published: (2024)
Small-to-Large Generalization: Data Influences Models Consistently Across Scale
by: Khaddaj, Alaa, et al.
Published: (2025)
by: Khaddaj, Alaa, et al.
Published: (2025)
Optimizing ML Training with Metagradient Descent
by: Engstrom, Logan, et al.
Published: (2025)
by: Engstrom, Logan, et al.
Published: (2025)
Decomposing and Editing Predictions by Modeling Model Computation
by: Shah, Harshay, et al.
Published: (2024)
by: Shah, Harshay, et al.
Published: (2024)
User Strategization and Trustworthy Algorithms
by: Cen, Sarah H., et al.
Published: (2023)
by: Cen, Sarah H., et al.
Published: (2023)
Everything You Always Wanted to Know About Storage Compressibility of Pre-Trained ML Models but Were Afraid to Ask
by: Su, Zhaoyuan, et al.
Published: (2024)
by: Su, Zhaoyuan, et al.
Published: (2024)
Why Ask One When You Can Ask $k$? Learning-to-Defer to the Top-$k$ Experts
by: Montreuil, Yannis, et al.
Published: (2025)
by: Montreuil, Yannis, et al.
Published: (2025)
Ask a Strong LLM Judge when Your Reward Model is Uncertain
by: Xu, Zhenghao, et al.
Published: (2025)
by: Xu, Zhenghao, et al.
Published: (2025)
Data Debiasing with Datamodels (D3M): Improving Subgroup Robustness via Data Selection
by: Jain, Saachi, et al.
Published: (2024)
by: Jain, Saachi, et al.
Published: (2024)
DataMIL: Selecting Data for Robot Imitation Learning with Datamodels
by: Dass, Shivin, et al.
Published: (2025)
by: Dass, Shivin, et al.
Published: (2025)
AI Supply Chains: An Emerging Ecosystem of AI Actors, Products, and Services
by: Hopkins, Aspen, et al.
Published: (2025)
by: Hopkins, Aspen, et al.
Published: (2025)
Attribute-to-Delete: Machine Unlearning via Datamodel Matching
by: Georgiev, Kristian, et al.
Published: (2024)
by: Georgiev, Kristian, et al.
Published: (2024)
Adversarial Training Improves Generalization Under Distribution Shifts in Bioacoustics
by: Heinrich, René, et al.
Published: (2025)
by: Heinrich, René, et al.
Published: (2025)
Measuring Strategization in Recommendation: Users Adapt Their Behavior to Shape Future Content
by: Cen, Sarah H., et al.
Published: (2024)
by: Cen, Sarah H., et al.
Published: (2024)
Large-Scale, Longitudinal Study of Large Language Models During the 2024 US Election Season
by: Cen, Sarah H., et al.
Published: (2025)
by: Cen, Sarah H., et al.
Published: (2025)
Know When You're Wrong: Aligning Confidence with Correctness for LLM Error Detection
by: Xiaohu, Xie, et al.
Published: (2026)
by: Xiaohu, Xie, et al.
Published: (2026)
Can You Trust an LLM with Your Life-Changing Decision? An Investigation into AI High-Stakes Responses
by: Cahyono, Joshua Adrian, et al.
Published: (2025)
by: Cahyono, Joshua Adrian, et al.
Published: (2025)
To Ask or Not to Ask: Learning to Require Human Feedback
by: Pugnana, Andrea, et al.
Published: (2025)
by: Pugnana, Andrea, et al.
Published: (2025)
Layer of Truth: Probing Belief Shifts under Continual Pre-Training Poisoning
by: Churina, Svetlana, et al.
Published: (2025)
by: Churina, Svetlana, et al.
Published: (2025)
Does Your Optimizer Care How You Normalize? Normalization-Optimizer Coupling in LLM Training
by: Abouzeid, Abdelrahman
Published: (2026)
by: Abouzeid, Abdelrahman
Published: (2026)
Can Your Generative Model Detect Out-of-Distribution Covariate Shift?
by: Viviers, Christiaan, et al.
Published: (2024)
by: Viviers, Christiaan, et al.
Published: (2024)
When Prompts Interact: Assessing Prompt Arithmetic for Deconfounding under Distribution Shift
by: Sheng, Zhecheng, et al.
Published: (2026)
by: Sheng, Zhecheng, et al.
Published: (2026)
Assessing the Impact of Distribution Shift on Reinforcement Learning Performance
by: Fujimoto, Ted, et al.
Published: (2024)
by: Fujimoto, Ted, et al.
Published: (2024)
Too Big to Think: Capacity, Memorization, and Generalization in Pre-Trained Transformers
by: Barron, Joshua, et al.
Published: (2025)
by: Barron, Joshua, et al.
Published: (2025)
Assessing Pre-Trained Models for Transfer Learning Through Distribution of Spectral Components
by: Zhang, Tengxue, et al.
Published: (2024)
by: Zhang, Tengxue, et al.
Published: (2024)
When and What to Ask: AskBench and Rubric-Guided RLVR for LLM Clarification
by: Zhao, Jiale, et al.
Published: (2026)
by: Zhao, Jiale, et al.
Published: (2026)
Benchmarking Distribution Shift in Tabular Data with TableShift
by: Gardner, Josh, et al.
Published: (2023)
by: Gardner, Josh, et al.
Published: (2023)
Watch Your Steps: Observable and Modular Chains of Thought
by: Cohen, Cassandra A., et al.
Published: (2024)
by: Cohen, Cassandra A., et al.
Published: (2024)
Distribution Shift Aware Neural Tabular Learning
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
Believe Your Model: Distribution-Guided Confidence Calibration
by: Yang, Xizhong, et al.
Published: (2026)
by: Yang, Xizhong, et al.
Published: (2026)
Your Pre-trained LLM is Secretly an Unsupervised Confidence Calibrator
by: Luo, Beier, et al.
Published: (2025)
by: Luo, Beier, et al.
Published: (2025)
ShiftKD: Benchmarking Knowledge Distillation under Distribution Shift
by: Zhang, Songming, et al.
Published: (2023)
by: Zhang, Songming, et al.
Published: (2023)
Distributed and Decentralised Training: Technical Governance Challenges in a Shifting AI Landscape
by: Kryś, Jakub, et al.
Published: (2025)
by: Kryś, Jakub, et al.
Published: (2025)
CTRL Your Shift: Clustered Transfer Residual Learning for Many Small Datasets
by: Jain, Gauri, et al.
Published: (2025)
by: Jain, Gauri, et al.
Published: (2025)
From Passive to Active Reasoning: Can Large Language Models Ask the Right Questions under Incomplete Information?
by: Zhou, Zhanke, et al.
Published: (2025)
by: Zhou, Zhanke, et al.
Published: (2025)
Graphs Generalization under Distribution Shifts
by: Tian, Qin, et al.
Published: (2024)
by: Tian, Qin, et al.
Published: (2024)
What Would You Ask When You First Saw $a^2+b^2=c^2$? Evaluating LLM on Curiosity-Driven Questioning
by: Javaji, Shashidhar Reddy, et al.
Published: (2024)
by: Javaji, Shashidhar Reddy, et al.
Published: (2024)
Similar Items
-
Do Large Language Model Benchmarks Test Reliability?
by: Vendrow, Joshua, et al.
Published: (2025) -
Learning to Attribute with Attention
by: Cohen-Wang, Benjamin, et al.
Published: (2025) -
ContextCite: Attributing Model Generation to Context
by: Cohen-Wang, Benjamin, et al.
Published: (2024) -
DsDm: Model-Aware Dataset Selection with Datamodels
by: Engstrom, Logan, et al.
Published: (2024) -
Small-to-Large Generalization: Data Influences Models Consistently Across Scale
by: Khaddaj, Alaa, et al.
Published: (2025)