All models are wrong, some are useful: Model Selection with Limited Labels
Fuente:
arXiv
Saved in:
| Main Authors: | Okanovic, Patrik, Kirsch, Andreas, Kasper, Jannes, Hoefler, Torsten, Krause, Andreas, Gürel, Nezihe Merve |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Large Language Model Selection with Limited Annotations
by: Durmazkeser, Yavuz, et al.
Published: (2026)
by: Durmazkeser, Yavuz, et al.
Published: (2026)
Active Model Selection for Large Language Models
by: Durmazkeser, Yavuz, et al.
Published: (2025)
by: Durmazkeser, Yavuz, et al.
Published: (2025)
EntryPrune: Neural Network Feature Selection using First Impressions
by: Zimmer, Felix, et al.
Published: (2024)
by: Zimmer, Felix, et al.
Published: (2024)
Evaluating the Limits of Large Language Models in Multilingual Legal Reasoning
by: Ioannou, Antreas, et al.
Published: (2025)
by: Ioannou, Antreas, et al.
Published: (2025)
Confounder Detection via Treatment Intent: A New Observational Study Design
by: Plecko, Drago, et al.
Published: (2026)
by: Plecko, Drago, et al.
Published: (2026)
Epidemiology of Large Language Models: A Benchmark for Observational Distribution Knowledge
by: Plecko, Drago, et al.
Published: (2025)
by: Plecko, Drago, et al.
Published: (2025)
Collaboratively Learning Federated Models from Noisy Decentralized Data
by: Li, Haoyuan, et al.
Published: (2024)
by: Li, Haoyuan, et al.
Published: (2024)
COLEP: Certifiably Robust Learning-Reasoning Conformal Prediction via Probabilistic Circuits
by: Kang, Mintong, et al.
Published: (2024)
by: Kang, Mintong, et al.
Published: (2024)
Advancing Deep Active Learning & Data Subset Selection: Unifying Principles with Information-Theory Intuitions
by: Kirsch, Andreas
Published: (2024)
by: Kirsch, Andreas
Published: (2024)
(Implicit) Ensembles of Ensembles: Epistemic Uncertainty Collapse in Large Models
by: Kirsch, Andreas
Published: (2024)
by: Kirsch, Andreas
Published: (2024)
When Data Is Scarce: Scaling Sparse Language Models with Repeated Training
by: Wu, Boqian, et al.
Published: (2026)
by: Wu, Boqian, et al.
Published: (2026)
Memory-Efficient LLM Training with Dynamic Sparsity: From Stability to Practical Scaling
by: Xiao, Qiao, et al.
Published: (2026)
by: Xiao, Qiao, et al.
Published: (2026)
BLaST: High Performance Inference and Pretraining using BLock Sparse Transformers
by: Okanovic, Patrik, et al.
Published: (2025)
by: Okanovic, Patrik, et al.
Published: (2025)
Scaling Laws of Global Weather Models
by: Yu, Yuejiang, et al.
Published: (2026)
by: Yu, Yuejiang, et al.
Published: (2026)
Specialization after Generalization: Towards Understanding Test-Time Training in Foundation Models
by: Hübotter, Jonas, et al.
Published: (2025)
by: Hübotter, Jonas, et al.
Published: (2025)
Causal Modeling with Stationary Diffusions
by: Lorch, Lars, et al.
Published: (2023)
by: Lorch, Lars, et al.
Published: (2023)
The Benefits and Risks of Transductive Approaches for AI Fairness
by: Razzak, Muhammed, et al.
Published: (2024)
by: Razzak, Muhammed, et al.
Published: (2024)
Learning Safety Constraints for Large Language Models
by: Chen, Xin, et al.
Published: (2025)
by: Chen, Xin, et al.
Published: (2025)
Safe Exploration Using Bayesian World Models and Log-Barrier Optimization
by: As, Yarden, et al.
Published: (2024)
by: As, Yarden, et al.
Published: (2024)
When three experiments are better than two: Avoiding intractable correlated aleatoric uncertainty by leveraging a novel bias--variance tradeoff
by: Scherer, Paul, et al.
Published: (2025)
by: Scherer, Paul, et al.
Published: (2025)
Near-Optimal Sparse Allreduce for Distributed Deep Learning
by: Li, Shigang, et al.
Published: (2022)
by: Li, Shigang, et al.
Published: (2022)
Chimera: Efficiently Training Large-Scale Neural Networks with Bidirectional Pipelines
by: Li, Shigang, et al.
Published: (2021)
by: Li, Shigang, et al.
Published: (2021)
MARLIN: Mixed-Precision Auto-Regressive Parallel Inference on Large Language Models
by: Frantar, Elias, et al.
Published: (2024)
by: Frantar, Elias, et al.
Published: (2024)
Probabilistic Artificial Intelligence
by: Krause, Andreas, et al.
Published: (2025)
by: Krause, Andreas, et al.
Published: (2025)
Optimistic Task Inference for Behavior Foundation Models
by: Rupf, Thomas, et al.
Published: (2025)
by: Rupf, Thomas, et al.
Published: (2025)
Grid Games: The Power of Multiple Grids for Quantizing Large Language Models
by: Egiazarian, Vage, et al.
Published: (2026)
by: Egiazarian, Vage, et al.
Published: (2026)
Uncertainty-Aware Robotic World Model Makes Offline Model-Based Reinforcement Learning Work on Real Robots
by: Li, Chenhao, et al.
Published: (2025)
by: Li, Chenhao, et al.
Published: (2025)
C-RAG: Certified Generation Risks for Retrieval-Augmented Language Models
by: Kang, Mintong, et al.
Published: (2024)
by: Kang, Mintong, et al.
Published: (2024)
Finetuning a Weather Foundation Model with Lightweight Decoders for Unseen Physical Processes
by: Lehmann, Fanny, et al.
Published: (2025)
by: Lehmann, Fanny, et al.
Published: (2025)
Exploring Design Choices for Autoregressive Deep Learning Climate Models
by: Gallusser, Florian, et al.
Published: (2025)
by: Gallusser, Florian, et al.
Published: (2025)
Directed Exploration in Reinforcement Learning from Linear Temporal Logic
by: Bagatella, Marco, et al.
Published: (2024)
by: Bagatella, Marco, et al.
Published: (2024)
SHAP values via sparse Fourier representation
by: Gorji, Ali, et al.
Published: (2024)
by: Gorji, Ali, et al.
Published: (2024)
Residual Deep Gaussian Processes on Manifolds
by: Wyrwal, Kacper, et al.
Published: (2024)
by: Wyrwal, Kacper, et al.
Published: (2024)
Towards Understanding and Avoiding Limitations of Convolutions on Graphs
by: Roth, Andreas
Published: (2026)
by: Roth, Andreas
Published: (2026)
EfQAT: An Efficient Framework for Quantization-Aware Training
by: Ashkboos, Saleh, et al.
Published: (2024)
by: Ashkboos, Saleh, et al.
Published: (2024)
Beyond Outliers: A Study of Optimizers Under Quantization
by: Vlassis, Georgios, et al.
Published: (2025)
by: Vlassis, Georgios, et al.
Published: (2025)
Robotic World Model: A Neural Network Simulator for Robust Policy Optimization in Robotics
by: Li, Chenhao, et al.
Published: (2025)
by: Li, Chenhao, et al.
Published: (2025)
High Performance Unstructured SpMM Computation Using Tensor Cores
by: Okanovic, Patrik, et al.
Published: (2024)
by: Okanovic, Patrik, et al.
Published: (2024)
DaCe AD: Unifying High-Performance Automatic Differentiation for Machine Learning and Scientific Computing
by: Boudaoud, Afif, et al.
Published: (2025)
by: Boudaoud, Afif, et al.
Published: (2025)
Generative Intervention Models for Causal Perturbation Modeling
by: Schneider, Nora, et al.
Published: (2024)
by: Schneider, Nora, et al.
Published: (2024)
Similar Items
-
Large Language Model Selection with Limited Annotations
by: Durmazkeser, Yavuz, et al.
Published: (2026) -
Active Model Selection for Large Language Models
by: Durmazkeser, Yavuz, et al.
Published: (2025) -
EntryPrune: Neural Network Feature Selection using First Impressions
by: Zimmer, Felix, et al.
Published: (2024) -
Evaluating the Limits of Large Language Models in Multilingual Legal Reasoning
by: Ioannou, Antreas, et al.
Published: (2025) -
Confounder Detection via Treatment Intent: A New Observational Study Design
by: Plecko, Drago, et al.
Published: (2026)