Is your model predicting the past?
Fuente:
arXiv
Saved in:
| Main Authors: | Hardt, Moritz, Kim, Michael P. |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ImageNot: A contrast with ImageNet preserves model rankings
by: Salaudeen, Olawale, et al.
Published: (2024)
by: Salaudeen, Olawale, et al.
Published: (2024)
Inherent Trade-Offs between Diversity and Stability in Multi-Task Benchmarks
by: Zhang, Guanhua, et al.
Published: (2024)
by: Zhang, Guanhua, et al.
Published: (2024)
Good Allocations from Bad Estimates
by: Casacuberta, Sílvia, et al.
Published: (2026)
by: Casacuberta, Sílvia, et al.
Published: (2026)
Test-Time Training on Nearest Neighbors for Large Language Models
by: Hardt, Moritz, et al.
Published: (2023)
by: Hardt, Moritz, et al.
Published: (2023)
Performative Prediction: Past and Future
by: Hardt, Moritz, et al.
Published: (2023)
by: Hardt, Moritz, et al.
Published: (2023)
Don't Label Twice: Quantity Beats Quality when Comparing Binary Classifiers on a Budget
by: Dorner, Florian E., et al.
Published: (2024)
by: Dorner, Florian E., et al.
Published: (2024)
Do causal predictors generalize better to new domains?
by: Nastl, Vivian Y., et al.
Published: (2024)
by: Nastl, Vivian Y., et al.
Published: (2024)
What Makes ImageNet Look Unlike LAION
by: Shirali, Ali, et al.
Published: (2023)
by: Shirali, Ali, et al.
Published: (2023)
Evaluating language models as risk scores
by: Cruz, André F., et al.
Published: (2024)
by: Cruz, André F., et al.
Published: (2024)
Unprocessing Seven Years of Algorithmic Fairness
by: Cruz, André F., et al.
Published: (2023)
by: Cruz, André F., et al.
Published: (2023)
How Benchmark Prediction from Fewer Data Misses the Mark
by: Zhang, Guanhua, et al.
Published: (2025)
by: Zhang, Guanhua, et al.
Published: (2025)
Computational Arbitrage in AI Model Markets
by: Olmedo, Ricardo, et al.
Published: (2026)
by: Olmedo, Ricardo, et al.
Published: (2026)
Limits to scalable evaluation at the frontier: LLM as Judge won't beat twice the data
by: Dorner, Florian E., et al.
Published: (2024)
by: Dorner, Florian E., et al.
Published: (2024)
Train-before-Test Harmonizes Language Model Rankings
by: Zhang, Guanhua, et al.
Published: (2025)
by: Zhang, Guanhua, et al.
Published: (2025)
Allocation Requires Prediction Only if Inequality Is Low
by: Shirali, Ali, et al.
Published: (2024)
by: Shirali, Ali, et al.
Published: (2024)
Leaderboard Incentives: Model Rankings under Strategic Post-Training
by: Chen, Yatong, et al.
Published: (2026)
by: Chen, Yatong, et al.
Published: (2026)
Limits to Predicting Online Speech Using Large Language Models
by: Remeli, Mina, et al.
Published: (2024)
by: Remeli, Mina, et al.
Published: (2024)
Training on the Test Task Confounds Evaluation and Emergence
by: Dominguez-Olmedo, Ricardo, et al.
Published: (2024)
by: Dominguez-Olmedo, Ricardo, et al.
Published: (2024)
Scaling Open-Ended Reasoning to Predict the Future
by: Chandak, Nikhil, et al.
Published: (2025)
by: Chandak, Nikhil, et al.
Published: (2025)
Algorithmic Collective Action in Machine Learning
by: Hardt, Moritz, et al.
Published: (2023)
by: Hardt, Moritz, et al.
Published: (2023)
Learning on the Job: Test-Time Curricula for Targeted Reinforcement Learning
by: Hübotter, Jonas, et al.
Published: (2025)
by: Hübotter, Jonas, et al.
Published: (2025)
Answer Matching Outperforms Multiple Choice for Language Model Evaluation
by: Chandak, Nikhil, et al.
Published: (2025)
by: Chandak, Nikhil, et al.
Published: (2025)
What happens to diffusion model likelihood when your model is conditional?
by: Cross, Mattias, et al.
Published: (2024)
by: Cross, Mattias, et al.
Published: (2024)
Is your algorithm unlearning or untraining?
by: Triantafillou, Eleni, et al.
Published: (2026)
by: Triantafillou, Eleni, et al.
Published: (2026)
On Surjectivity of Neural Networks: Can you elicit any behavior from your model?
by: Jiang, Haozhe, et al.
Published: (2025)
by: Jiang, Haozhe, et al.
Published: (2025)
How to make the most of your masked language model for protein engineering
by: McCarter, Calvin, et al.
Published: (2026)
by: McCarter, Calvin, et al.
Published: (2026)
Transformers for molecular property prediction: Lessons learned from the past five years
by: Sultan, Afnan, et al.
Published: (2024)
by: Sultan, Afnan, et al.
Published: (2024)
Data-driven development of cycle prediction models for lithium metal batteries using multi modal mining
by: Lee, Jaewoong, et al.
Published: (2024)
by: Lee, Jaewoong, et al.
Published: (2024)
Is your batch size the problem? Revisiting the Adam-SGD gap in language modeling
by: Srećković, Teodora, et al.
Published: (2025)
by: Srećković, Teodora, et al.
Published: (2025)
Toward Falsifying Causal Graphs Using a Permutation-Based Test
by: Eulig, Elias, et al.
Published: (2023)
by: Eulig, Elias, et al.
Published: (2023)
DEQuify your force field: More efficient simulations using deep equilibrium models
by: Burger, Andreas, et al.
Published: (2025)
by: Burger, Andreas, et al.
Published: (2025)
Learning from the past: predicting critical transitions with machine learning trained on surrogates of historical data
by: Ma, Zhiqin, et al.
Published: (2024)
by: Ma, Zhiqin, et al.
Published: (2024)
Lawma: The Power of Specialization for Legal Annotation
by: Dominguez-Olmedo, Ricardo, et al.
Published: (2024)
by: Dominguez-Olmedo, Ricardo, et al.
Published: (2024)
Matrix factorization and prediction for high dimensional co-occurrence count data via shared parameter alternating zero inflated Gamma model
by: Kim, Taejoon, et al.
Published: (2024)
by: Kim, Taejoon, et al.
Published: (2024)
FutureSim: Replaying World Events to Evaluate Adaptive Agents
by: Goel, Shashwat, et al.
Published: (2026)
by: Goel, Shashwat, et al.
Published: (2026)
Keep your distance: learning dispersed embeddings on $\mathbb{S}_m$
by: Tokarchuk, Evgeniia, et al.
Published: (2025)
by: Tokarchuk, Evgeniia, et al.
Published: (2025)
Multi-site modelling and reconstruction of past extreme skew surges along the French Atlantic coast
by: Huet, Nathan, et al.
Published: (2025)
by: Huet, Nathan, et al.
Published: (2025)
Surrogate modeling for interpreting black-box LLMs in medical predictions
by: Han, Changho, et al.
Published: (2026)
by: Han, Changho, et al.
Published: (2026)
Turbo your multi-modal classification with contrastive learning
by: Zhang, Zhiyu, et al.
Published: (2024)
by: Zhang, Zhiyu, et al.
Published: (2024)
Secret mixtures of experts inside your LLM
by: Boix-Adsera, Enric
Published: (2025)
by: Boix-Adsera, Enric
Published: (2025)
Similar Items
-
ImageNot: A contrast with ImageNet preserves model rankings
by: Salaudeen, Olawale, et al.
Published: (2024) -
Inherent Trade-Offs between Diversity and Stability in Multi-Task Benchmarks
by: Zhang, Guanhua, et al.
Published: (2024) -
Good Allocations from Bad Estimates
by: Casacuberta, Sílvia, et al.
Published: (2026) -
Test-Time Training on Nearest Neighbors for Large Language Models
by: Hardt, Moritz, et al.
Published: (2023) -
Performative Prediction: Past and Future
by: Hardt, Moritz, et al.
Published: (2023)