A Guide to Robust Generalization: The Impact of Architecture, Pre-training, and Optimization Strategy
Fuente:
arXiv
Saved in:
| Main Authors: | Heuillet, Maxime, Bhagwatkar, Rishika, Ngnawé, Jonas, Pequignot, Yann, Larouche, Alexandre, Gagné, Christian, Rish, Irina, Ahmad, Ola, Durand, Audrey |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Robust Fine-Tuning from Non-Robust Pretrained Models: Mitigating Suboptimal Transfer With Epsilon-Scheduling
by: Ngnawé, Jonas, et al.
Published: (2025)
by: Ngnawé, Jonas, et al.
Published: (2025)
Randomized Confidence Bounds for Stochastic Partial Monitoring
by: Heuillet, Maxime, et al.
Published: (2024)
by: Heuillet, Maxime, et al.
Published: (2024)
Neural Active Learning Meets the Partial Monitoring Framework
by: Heuillet, Maxime, et al.
Published: (2024)
by: Heuillet, Maxime, et al.
Published: (2024)
On the Adversarial Robustness of Discrete Image Tokenizers
by: Bhagwatkar, Rishika, et al.
Published: (2026)
by: Bhagwatkar, Rishika, et al.
Published: (2026)
Detecting Brittle Decisions for Free: Leveraging Margin Consistency in Deep Robust Classifiers
by: Ngnawé, Jonas, et al.
Published: (2024)
by: Ngnawé, Jonas, et al.
Published: (2024)
LLM-as-a-Judge: Toward World Models for Slate Recommendation Systems
by: Bonin, Baptiste, et al.
Published: (2025)
by: Bonin, Baptiste, et al.
Published: (2025)
Signal from Structure: Exploiting Submodular Upper Bounds in Generative Flow Networks
by: Larouche, Alexandre, et al.
Published: (2026)
by: Larouche, Alexandre, et al.
Published: (2026)
A Layer Selection Approach to Test Time Adaptation
by: Sahoo, Sabyasachi, et al.
Published: (2024)
by: Sahoo, Sabyasachi, et al.
Published: (2024)
CAVE: Detecting and Explaining Commonsense Anomalies in Visual Environments
by: Bhagwatkar, Rishika, et al.
Published: (2025)
by: Bhagwatkar, Rishika, et al.
Published: (2025)
Towards Adversarially Robust Vision-Language Models: Insights from Design Choices and Prompt Formatting Techniques
by: Bhagwatkar, Rishika, et al.
Published: (2024)
by: Bhagwatkar, Rishika, et al.
Published: (2024)
Nested-ReFT: Efficient Reinforcement Learning for Large Language Model Fine-Tuning via Off-Policy Rollouts
by: Heuillet, Maxime, et al.
Published: (2025)
by: Heuillet, Maxime, et al.
Published: (2025)
Indirect Prompt Injections: Are Firewalls All You Need, or Stronger Benchmarks?
by: Bhagwatkar, Rishika, et al.
Published: (2025)
by: Bhagwatkar, Rishika, et al.
Published: (2025)
Robustmix: Improving Robustness by Regularizing the Frequency Bias of Deep Nets
by: Ngnawe, Jonas, et al.
Published: (2023)
by: Ngnawe, Jonas, et al.
Published: (2023)
Towards Better: A motivated introduction to better-quasi-orders
by: Pequignot, Yann
Published: (2016)
by: Pequignot, Yann
Published: (2016)
Finite versus infinite: an insufficient shift
by: Pequignot, Yann
Published: (2016)
by: Pequignot, Yann
Published: (2016)
TrackPGD: Efficient Adversarial Attack using Object Binary Masks against Robust Transformer Trackers
by: Nokabadi, Fatemeh Nourilenjan, et al.
Published: (2024)
by: Nokabadi, Fatemeh Nourilenjan, et al.
Published: (2024)
Simple and Scalable Strategies to Continually Pre-train Large Language Models
by: Ibrahim, Adam, et al.
Published: (2024)
by: Ibrahim, Adam, et al.
Published: (2024)
A well-quasi-order for continuous functions
by: Carroy, Raphaël, et al.
Published: (2024)
by: Carroy, Raphaël, et al.
Published: (2024)
Warming Up for Zeroth-Order Federated Pre-Training with Low Resource Clients
by: Legate, Gwen, et al.
Published: (2025)
by: Legate, Gwen, et al.
Published: (2025)
Beyond Cosine Decay: On the effectiveness of Infinite Learning Rate Schedule for Continual Pre-training
by: Singh, Vaibhav, et al.
Published: (2025)
by: Singh, Vaibhav, et al.
Published: (2025)
Embeddability on functions: order and chaos
by: Carroy, Raphaël, et al.
Published: (2018)
by: Carroy, Raphaël, et al.
Published: (2018)
Knowledge Distillation for Federated Learning: a Practical Guide
by: Mora, Alessio, et al.
Published: (2022)
by: Mora, Alessio, et al.
Published: (2022)
Partial Order in Chaos: Consensus on Feature Attributions in the Rashomon Set
by: Laberge, Gabriel, et al.
Published: (2021)
by: Laberge, Gabriel, et al.
Published: (2021)
Enhancing Context Through Contrast
by: Ambilduke, Kshitij, et al.
Published: (2024)
by: Ambilduke, Kshitij, et al.
Published: (2024)
Training a neural network to rapidly identify candidate gravitational-wave events in the lower mass gap
by: Raza, Nayyer, et al.
Published: (2026)
by: Raza, Nayyer, et al.
Published: (2026)
GWSkyNet-Multi II: an updated machine learning model for rapid classification of gravitational-wave events
by: Raza, Nayyer, et al.
Published: (2025)
by: Raza, Nayyer, et al.
Published: (2025)
MoRE: Batch-Robust Multi-Omics Representations from Frozen Pre-trained Transformers
by: Chen, Audrey Pei-Hsuan
Published: (2025)
by: Chen, Audrey Pei-Hsuan
Published: (2025)
RETRACTED: Optimizing Electron‐Beam Photothermal Pyrolysis Parameters for Enhanced Molecular Biodegradability of Polyvinyl Chloride Plastics
by: Rishika Porandla
Published: (2024)
by: Rishika Porandla
Published: (2024)
Preference-based learning for news headline recommendation
by: Bouras, Alexandre, et al.
Published: (2025)
by: Bouras, Alexandre, et al.
Published: (2025)
Continual Pre-training of MoEs: How robust is your router?
by: Thérien, Benjamin, et al.
Published: (2025)
by: Thérien, Benjamin, et al.
Published: (2025)
El oficio de pensar. Visiones y papeles de Pierre Bourdieu
by: Bruno Péquignot
Published: (2002)
by: Bruno Péquignot
Published: (2002)
Third-dredge-up oxygen in planetary nebulae?
by: D. Péquignot
Published: (2000)
by: D. Péquignot
Published: (2000)
La sociologie de l’art et de la culture en France: un état des lieux
by: Bruno Péquignot
Published: (2005)
by: Bruno Péquignot
Published: (2005)
Caught between two states: The compromise in acclimation of photosynthesis, transpiration and mesophyll conductance to different amplitudes of fluctuating irradiance
by: Maxime Durand, et al.
Published: (2024)
by: Maxime Durand, et al.
Published: (2024)
Towards ethical multimodal systems
by: Roger, Alexis, et al.
Published: (2023)
by: Roger, Alexis, et al.
Published: (2023)
A Quick Guide to Nearby Young Association
by: Gagné, Jonathan
Published: (2024)
by: Gagné, Jonathan
Published: (2024)
Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting
by: Rasul, Kashif, et al.
Published: (2023)
by: Rasul, Kashif, et al.
Published: (2023)
Corps-matière et jouissance: le rêve d'un nouveau corps
by: Catherine Desprats Péquignot
Published: (2008)
by: Catherine Desprats Péquignot
Published: (2008)
Corps-matière et jouissance: le rêve d’un nouveau corps
by: Catherine Desprats Péquignot
Published: (2008)
by: Catherine Desprats Péquignot
Published: (2008)
On the Fundamental Limitations of Dual Static CVaR Decompositions in Markov Decision Processes
by: Godbout, Mathieu, et al.
Published: (2025)
by: Godbout, Mathieu, et al.
Published: (2025)
Similar Items
-
Robust Fine-Tuning from Non-Robust Pretrained Models: Mitigating Suboptimal Transfer With Epsilon-Scheduling
by: Ngnawé, Jonas, et al.
Published: (2025) -
Randomized Confidence Bounds for Stochastic Partial Monitoring
by: Heuillet, Maxime, et al.
Published: (2024) -
Neural Active Learning Meets the Partial Monitoring Framework
by: Heuillet, Maxime, et al.
Published: (2024) -
On the Adversarial Robustness of Discrete Image Tokenizers
by: Bhagwatkar, Rishika, et al.
Published: (2026) -
Detecting Brittle Decisions for Free: Leveraging Margin Consistency in Deep Robust Classifiers
by: Ngnawé, Jonas, et al.
Published: (2024)