Occam's model: Selecting simpler representations for better transferability estimation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Singh, Prabhant, Hess, Sibylle, Vanschoren, Joaquin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How NOT to benchmark your SITE metric: Beyond Static Leaderboards and Towards Realistic Evaluation
von: Singh, Prabhant, et al.
Veröffentlicht: (2025)
von: Singh, Prabhant, et al.
Veröffentlicht: (2025)
Meta-Learning for Unsupervised Outlier Detection with Optimal Transport
von: Singh, Prabhant, et al.
Veröffentlicht: (2022)
von: Singh, Prabhant, et al.
Veröffentlicht: (2022)
CLAMS: A System for Zero-Shot Model Selection for Clustering
von: Singh, Prabhant, et al.
Veröffentlicht: (2024)
von: Singh, Prabhant, et al.
Veröffentlicht: (2024)
On Supernet Transfer Learning for Effective Task Adaptation
von: Singh, Prabhant, et al.
Veröffentlicht: (2024)
von: Singh, Prabhant, et al.
Veröffentlicht: (2024)
Meta-Learning Transformers to Improve In-Context Generalization
von: Braccaioli, Lorenzo, et al.
Veröffentlicht: (2025)
von: Braccaioli, Lorenzo, et al.
Veröffentlicht: (2025)
Robustness of AutoML on Dirty Categorical Data
von: Bueno, Marcos L. P., et al.
Veröffentlicht: (2026)
von: Bueno, Marcos L. P., et al.
Veröffentlicht: (2026)
Automatic Combination of Sample Selection Strategies for Few-Shot Learning
von: Pecher, Branislav, et al.
Veröffentlicht: (2024)
von: Pecher, Branislav, et al.
Veröffentlicht: (2024)
In-context learning and Occam's razor
von: Elmoznino, Eric, et al.
Veröffentlicht: (2024)
von: Elmoznino, Eric, et al.
Veröffentlicht: (2024)
Deep neural networks have an inbuilt Occam's razor
von: Mingard, Chris, et al.
Veröffentlicht: (2023)
von: Mingard, Chris, et al.
Veröffentlicht: (2023)
Occam's Razor for Self Supervised Learning: What is Sufficient to Learn Good Representations?
von: Ibrahim, Mark, et al.
Veröffentlicht: (2024)
von: Ibrahim, Mark, et al.
Veröffentlicht: (2024)
Automated Machine Learning for Unsupervised Tabular Tasks
von: Singh, Prabhant, et al.
Veröffentlicht: (2025)
von: Singh, Prabhant, et al.
Veröffentlicht: (2025)
Automated Reinforcement Learning: An Overview
von: Afshar, Reza Refaei, et al.
Veröffentlicht: (2022)
von: Afshar, Reza Refaei, et al.
Veröffentlicht: (2022)
In-Context Occam's Razor: How Transformers Prefer Simpler Hypotheses on the Fly
von: Deora, Puneesh, et al.
Veröffentlicht: (2025)
von: Deora, Puneesh, et al.
Veröffentlicht: (2025)
OccamLLM: Fast and Exact Language Model Arithmetic in a Single Step
von: Dugan, Owen, et al.
Veröffentlicht: (2024)
von: Dugan, Owen, et al.
Veröffentlicht: (2024)
SemPPL: Predicting pseudo-labels for better contrastive representations
von: Bošnjak, Matko, et al.
Veröffentlicht: (2023)
von: Bošnjak, Matko, et al.
Veröffentlicht: (2023)
On the transferability of Sparse Autoencoders for interpreting compressed models
von: Gupte, Suchit, et al.
Veröffentlicht: (2025)
von: Gupte, Suchit, et al.
Veröffentlicht: (2025)
SemiOccam: A Robust Semi-Supervised Image Recognition Network Using Sparse Labels
von: Yann, Rui, et al.
Veröffentlicht: (2025)
von: Yann, Rui, et al.
Veröffentlicht: (2025)
Self-Improving Pretraining: using post-trained models to pretrain better models
von: Tan, Ellen Xiaoqing, et al.
Veröffentlicht: (2026)
von: Tan, Ellen Xiaoqing, et al.
Veröffentlicht: (2026)
Latent label distribution grid representation for modeling uncertainty
von: Sun, ShuNing, et al.
Veröffentlicht: (2025)
von: Sun, ShuNing, et al.
Veröffentlicht: (2025)
Language models are better than humans at next-token prediction
von: Shlegeris, Buck, et al.
Veröffentlicht: (2022)
von: Shlegeris, Buck, et al.
Veröffentlicht: (2022)
On-site estimation of battery electrochemical parameters via transfer learning based physics-informed neural network approach
von: Yeregui, Josu, et al.
Veröffentlicht: (2025)
von: Yeregui, Josu, et al.
Veröffentlicht: (2025)
$μ$pscaling small models: Principled warm starts and hyperparameter transfer
von: Ma, Yuxin, et al.
Veröffentlicht: (2026)
von: Ma, Yuxin, et al.
Veröffentlicht: (2026)
Simmering: Sufficient is better than optimal for training neural networks
von: Babayan, Irina, et al.
Veröffentlicht: (2024)
von: Babayan, Irina, et al.
Veröffentlicht: (2024)
Optimal rates for density and mode estimation with expand-and-sparsify representations
von: Sinha, Kaushik, et al.
Veröffentlicht: (2026)
von: Sinha, Kaushik, et al.
Veröffentlicht: (2026)
Bilinear representation mitigates reversal curse and enables consistent model editing
von: Kim, Dong-Kyum, et al.
Veröffentlicht: (2025)
von: Kim, Dong-Kyum, et al.
Veröffentlicht: (2025)
Sparse deepfake detection promotes better disentanglement
von: Teissier, Antoine, et al.
Veröffentlicht: (2025)
von: Teissier, Antoine, et al.
Veröffentlicht: (2025)
An advantage based policy transfer algorithm for reinforcement learning with measures of transferability
von: Alam, Md Ferdous, et al.
Veröffentlicht: (2023)
von: Alam, Md Ferdous, et al.
Veröffentlicht: (2023)
Carré du champ flow matching: better quality-generalisation tradeoff in generative models
von: Bamberger, Jacob, et al.
Veröffentlicht: (2025)
von: Bamberger, Jacob, et al.
Veröffentlicht: (2025)
Can We Understand Plasticity Through Neural Collapse?
von: Bonifazi, Guglielmo, et al.
Veröffentlicht: (2024)
von: Bonifazi, Guglielmo, et al.
Veröffentlicht: (2024)
When narrower is better: the narrow width limit of Bayesian parallel branching neural networks
von: Zhang, Zechen, et al.
Veröffentlicht: (2024)
von: Zhang, Zechen, et al.
Veröffentlicht: (2024)
GeckOpt: LLM System Efficiency via Intent-Based Tool Selection
von: Fore, Michael, et al.
Veröffentlicht: (2024)
von: Fore, Michael, et al.
Veröffentlicht: (2024)
Constrained latent state modeling: A unifying perspective on representation learning under competing constraints
von: Quellec, Gwenolé
Veröffentlicht: (2026)
von: Quellec, Gwenolé
Veröffentlicht: (2026)
LaTiM: Longitudinal representation learning in continuous-time models to predict disease progression
von: Zeghlache, Rachid, et al.
Veröffentlicht: (2024)
von: Zeghlache, Rachid, et al.
Veröffentlicht: (2024)
What is causal about causal models and representations?
von: Jørgensen, Frederik Hytting, et al.
Veröffentlicht: (2025)
von: Jørgensen, Frederik Hytting, et al.
Veröffentlicht: (2025)
FLAG: Foundation model representation with Latent diffusion Alignment via Graph for spatial gene expression prediction
von: Si, Qi, et al.
Veröffentlicht: (2026)
von: Si, Qi, et al.
Veröffentlicht: (2026)
Gauge-invariant representation holonomy
von: Sevetlidis, Vasileios, et al.
Veröffentlicht: (2026)
von: Sevetlidis, Vasileios, et al.
Veröffentlicht: (2026)
NODE-AdvGAN: Improving the transferability and perceptual similarity of adversarial examples by dynamic-system-driven adversarial generative model
von: Xie, Xinheng, et al.
Veröffentlicht: (2024)
von: Xie, Xinheng, et al.
Veröffentlicht: (2024)
ExplainerPFN: Towards tabular foundation models for model-free zero-shot feature importance estimations
von: Fonseca, Joao, et al.
Veröffentlicht: (2026)
von: Fonseca, Joao, et al.
Veröffentlicht: (2026)
Scoring rule nets: beyond mean target prediction in multivariate regression
von: Roordink, Daan, et al.
Veröffentlicht: (2024)
von: Roordink, Daan, et al.
Veröffentlicht: (2024)
Search-contempt: a hybrid MCTS algorithm for training AlphaZero-like engines with better computational efficiency
von: Joshi, Ameya
Veröffentlicht: (2025)
von: Joshi, Ameya
Veröffentlicht: (2025)
Ähnliche Einträge
-
How NOT to benchmark your SITE metric: Beyond Static Leaderboards and Towards Realistic Evaluation
von: Singh, Prabhant, et al.
Veröffentlicht: (2025) -
Meta-Learning for Unsupervised Outlier Detection with Optimal Transport
von: Singh, Prabhant, et al.
Veröffentlicht: (2022) -
CLAMS: A System for Zero-Shot Model Selection for Clustering
von: Singh, Prabhant, et al.
Veröffentlicht: (2024) -
On Supernet Transfer Learning for Effective Task Adaptation
von: Singh, Prabhant, et al.
Veröffentlicht: (2024) -
Meta-Learning Transformers to Improve In-Context Generalization
von: Braccaioli, Lorenzo, et al.
Veröffentlicht: (2025)