Fast Generalization after Interpolation via Critically Damped Momentum Optimization
Fuente:
arXiv
Salvato in:
| Autori principali: | Muscarnera, Luca, Estévez, Silas Ruhrberg, Xiao, Yuanzhang, Van der Schaar, Mihaela |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Knowledge-Informed Kernel State Reconstruction from Heterogeneous Partial Observations
di: Muscarnera, Luca, et al.
Pubblicazione: (2026)
di: Muscarnera, Luca, et al.
Pubblicazione: (2026)
Automatic Construction of Clinical Scoring Systems with LLM Agents
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2026)
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2026)
Hypothesis Hunting with Evolving Networks of Autonomous Scientific Agents
di: Liu, Tennison, et al.
Pubblicazione: (2025)
di: Liu, Tennison, et al.
Pubblicazione: (2025)
Autoformulation of Mathematical Optimization Models Using LLMs
di: Astorga, Nicolás, et al.
Pubblicazione: (2024)
di: Astorga, Nicolás, et al.
Pubblicazione: (2024)
Timely Clinical Diagnosis through Active Test Selection
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2025)
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2025)
CellBRIDGE: Learning Cellular Trajectories via Interaction-Aware Alignment
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2026)
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2026)
Deep Hierarchical Learning with Nested Subspace Networks for Large Language Models
di: Rauba, Paulius, et al.
Pubblicazione: (2025)
di: Rauba, Paulius, et al.
Pubblicazione: (2025)
Shape Arithmetic Expressions: Advancing Scientific Discovery Beyond Closed-Form Equations
di: Kacprzyk, Krzysztof, et al.
Pubblicazione: (2024)
di: Kacprzyk, Krzysztof, et al.
Pubblicazione: (2024)
No Equations Needed: Learning System Dynamics Without Relying on Closed-Form ODEs
di: Kacprzyk, Krzysztof, et al.
Pubblicazione: (2025)
di: Kacprzyk, Krzysztof, et al.
Pubblicazione: (2025)
Towards Automated Knowledge Integration From Human-Interpretable Representations
di: Kobalczyk, Katarzyna, et al.
Pubblicazione: (2024)
di: Kobalczyk, Katarzyna, et al.
Pubblicazione: (2024)
Matchmaker: Self-Improving Large Language Model Programs for Schema Matching
di: Seedat, Nabeel, et al.
Pubblicazione: (2024)
di: Seedat, Nabeel, et al.
Pubblicazione: (2024)
Deep Learning for Motion Classification in Ankle Exoskeletons Using Surface EMG and IMU Signals
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2024)
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2024)
Decision Tree Induction Through LLMs via Semantically-Aware Evolution
di: Liu, Tennison, et al.
Pubblicazione: (2025)
di: Liu, Tennison, et al.
Pubblicazione: (2025)
Risk-Sensitive Diffusion: Robustly Optimizing Diffusion Models with Noisy Samples
di: Li, Yangming, et al.
Pubblicazione: (2024)
di: Li, Yangming, et al.
Pubblicazione: (2024)
Active Timepoint Selection for Learning Measure-Valued Trajectories
di: Huynh, Nicolas, et al.
Pubblicazione: (2026)
di: Huynh, Nicolas, et al.
Pubblicazione: (2026)
Hyperparameter Trajectory Inference with Conditional Lagrangian Optimal Transport
di: Amad, Harry, et al.
Pubblicazione: (2026)
di: Amad, Harry, et al.
Pubblicazione: (2026)
Not All Explanations for Deep Learning Phenomena Are Equally Valuable
di: Jeffares, Alan, et al.
Pubblicazione: (2025)
di: Jeffares, Alan, et al.
Pubblicazione: (2025)
Preference Learning for AI Alignment: a Causal Perspective
di: Kobalczyk, Katarzyna, et al.
Pubblicazione: (2025)
di: Kobalczyk, Katarzyna, et al.
Pubblicazione: (2025)
Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning
di: Sun, Hao, et al.
Pubblicazione: (2024)
di: Sun, Hao, et al.
Pubblicazione: (2024)
Position: All Current Generative Fidelity and Diversity Metrics are Flawed
di: Räisä, Ossi, et al.
Pubblicazione: (2025)
di: Räisä, Ossi, et al.
Pubblicazione: (2025)
Why Tabular Foundation Models Should Be a Research Priority
di: van Breugel, Boris, et al.
Pubblicazione: (2024)
di: van Breugel, Boris, et al.
Pubblicazione: (2024)
Query-Dependent Prompt Evaluation and Optimization with Offline Inverse RL
di: Sun, Hao, et al.
Pubblicazione: (2023)
di: Sun, Hao, et al.
Pubblicazione: (2023)
Discovery of Hidden Miscalibration Regimes
di: Kobalczyk, Katarzyna, et al.
Pubblicazione: (2026)
di: Kobalczyk, Katarzyna, et al.
Pubblicazione: (2026)
Language Bottleneck Models for Qualitative Knowledge State Modeling
di: Berthon, Antonin, et al.
Pubblicazione: (2025)
di: Berthon, Antonin, et al.
Pubblicazione: (2025)
On Error Propagation of Diffusion Models
di: Li, Yangming, et al.
Pubblicazione: (2023)
di: Li, Yangming, et al.
Pubblicazione: (2023)
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities
di: Sun, Hao, et al.
Pubblicazione: (2025)
di: Sun, Hao, et al.
Pubblicazione: (2025)
Large Language Models to Enhance Bayesian Optimization
di: Liu, Tennison, et al.
Pubblicazione: (2024)
di: Liu, Tennison, et al.
Pubblicazione: (2024)
Tiny Autoregressive Recursive Models
di: Rauba, Paulius, et al.
Pubblicazione: (2026)
di: Rauba, Paulius, et al.
Pubblicazione: (2026)
Interpretable Reward Modeling with Active Concept Bottlenecks
di: Laguna, Sonia, et al.
Pubblicazione: (2025)
di: Laguna, Sonia, et al.
Pubblicazione: (2025)
A Study of Posterior Stability for Time-Series Latent Diffusion
di: Li, Yangming, et al.
Pubblicazione: (2024)
di: Li, Yangming, et al.
Pubblicazione: (2024)
Automatically Learning Hybrid Digital Twins of Dynamical Systems
di: Holt, Samuel, et al.
Pubblicazione: (2024)
di: Holt, Samuel, et al.
Pubblicazione: (2024)
Why do Random Forests Work? Understanding Tree Ensembles as Self-Regularizing Adaptive Smoothers
di: Curth, Alicia, et al.
Pubblicazione: (2024)
di: Curth, Alicia, et al.
Pubblicazione: (2024)
Dissecting Sample Hardness: A Fine-Grained Analysis of Hardness Characterization Methods for Data-Centric AI
di: Seedat, Nabeel, et al.
Pubblicazione: (2024)
di: Seedat, Nabeel, et al.
Pubblicazione: (2024)
Quantifying Aleatoric Uncertainty of the Treatment Effect: A Novel Orthogonal Learner
di: Melnychuk, Valentyn, et al.
Pubblicazione: (2024)
di: Melnychuk, Valentyn, et al.
Pubblicazione: (2024)
Adaptive Experiment Design with Synthetic Controls
di: Hüyük, Alihan, et al.
Pubblicazione: (2024)
di: Hüyük, Alihan, et al.
Pubblicazione: (2024)
What's the next frontier for Data-centric AI? Data Savvy Agents
di: Seedat, Nabeel, et al.
Pubblicazione: (2025)
di: Seedat, Nabeel, et al.
Pubblicazione: (2025)
Deep Generative Symbolic Regression
di: Holt, Samuel, et al.
Pubblicazione: (2023)
di: Holt, Samuel, et al.
Pubblicazione: (2023)
Eliciting Numerical Predictive Distributions of LLMs Without Autoregression
di: Piskorz, Julianna, et al.
Pubblicazione: (2026)
di: Piskorz, Julianna, et al.
Pubblicazione: (2026)
Visualizing token importance for black-box language models
di: Rauba, Paulius, et al.
Pubblicazione: (2025)
di: Rauba, Paulius, et al.
Pubblicazione: (2025)
Simulating Viva Voce Examinations to Evaluate Clinical Reasoning in Large Language Models
di: Chiu, Christopher, et al.
Pubblicazione: (2025)
di: Chiu, Christopher, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Knowledge-Informed Kernel State Reconstruction from Heterogeneous Partial Observations
di: Muscarnera, Luca, et al.
Pubblicazione: (2026) -
Automatic Construction of Clinical Scoring Systems with LLM Agents
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2026) -
Hypothesis Hunting with Evolving Networks of Autonomous Scientific Agents
di: Liu, Tennison, et al.
Pubblicazione: (2025) -
Autoformulation of Mathematical Optimization Models Using LLMs
di: Astorga, Nicolás, et al.
Pubblicazione: (2024) -
Timely Clinical Diagnosis through Active Test Selection
di: Estévez, Silas Ruhrberg, et al.
Pubblicazione: (2025)