All AI Models are Wrong, but Some are Optimal
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Anand, Akhil S, Sawant, Shambhuraj, Reinhardt, Dirk, Gros, Sebastien |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Closing the Sim2Real Performance Gap in RL
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
Economic Model Predictive Control as a Solution to Markov Decision Processes
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2024)
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2024)
Data-Driven Predictive Control and MPC: Do we achieve optimality?
von: Anand, Akhil S, et al.
Veröffentlicht: (2024)
von: Anand, Akhil S, et al.
Veröffentlicht: (2024)
Synthesis of Model Predictive Control and Reinforcement Learning: Survey and Classification
von: Reiter, Rudolf, et al.
Veröffentlicht: (2025)
von: Reiter, Rudolf, et al.
Veröffentlicht: (2025)
Direct transfer of optimized controllers to similar systems using dimensionless MPC
von: Hromatko, Josip Kir, et al.
Veröffentlicht: (2025)
von: Hromatko, Josip Kir, et al.
Veröffentlicht: (2025)
CORL: Reinforcement Learning of MILP Policies Solved via Branch and Bound
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
von: Anand, Akhil S, et al.
Veröffentlicht: (2025)
Differentiable Nonlinear Model Predictive Control
von: Frey, Jonathan, et al.
Veröffentlicht: (2025)
von: Frey, Jonathan, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning using Action Projection: Safeguard the Policy or the Environment?
von: Markgraf, Hannah, et al.
Veröffentlicht: (2025)
von: Markgraf, Hannah, et al.
Veröffentlicht: (2025)
Application of Soft Actor-Critic Algorithms in Optimizing Wastewater Treatment with Time Delays Integration
von: Mohammadi, Esmaeel, et al.
Veröffentlicht: (2024)
von: Mohammadi, Esmaeel, et al.
Veröffentlicht: (2024)
Are Retrials All You Need? Enhancing Large Language Model Reasoning Without Verbalized Feedback
von: Potamitis, Nearchos, et al.
Veröffentlicht: (2025)
von: Potamitis, Nearchos, et al.
Veröffentlicht: (2025)
Small transformer architectures for task switching
von: Gros, Claudius
Veröffentlicht: (2025)
von: Gros, Claudius
Veröffentlicht: (2025)
Reorganizing attention-space geometry with expressive attention
von: Gros, Claudius
Veröffentlicht: (2024)
von: Gros, Claudius
Veröffentlicht: (2024)
Large Language Model Predicts Above Normal All India Summer Monsoon Rainfall in 2024
von: Sharma, Ujjawal, et al.
Veröffentlicht: (2024)
von: Sharma, Ujjawal, et al.
Veröffentlicht: (2024)
Understanding Generative AI Content with Embedding Models
von: Vargas, Max, et al.
Veröffentlicht: (2024)
von: Vargas, Max, et al.
Veröffentlicht: (2024)
Thinking Wrong in Silence: Backdoor Attacks on Continuous Latent Reasoning
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
Stable but Wrong: When More Data Degrades Scientific Conclusions
von: Zhang, Zhipeng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhipeng, et al.
Veröffentlicht: (2026)
ARMOR: Adaptive Resilience Against Model Poisoning Attacks in Continual Federated Learning for Mobile Indoor Localization
von: Gufran, Danish, et al.
Veröffentlicht: (2026)
von: Gufran, Danish, et al.
Veröffentlicht: (2026)
CosmicFish-HRM: Adaptive Reasoning via Hierarchical Recurrent Mechanisms in Compact Language Models
von: Lakkapragada, Venkat Akhil
Veröffentlicht: (2026)
von: Lakkapragada, Venkat Akhil
Veröffentlicht: (2026)
Hybrid Intent-Aware Personalization with Machine Learning and RAG-Enabled Large Language Models for Financial Services Marketing
von: Shanivendra, Akhil Chandra
Veröffentlicht: (2026)
von: Shanivendra, Akhil Chandra
Veröffentlicht: (2026)
When LLMs Learn to Be Consistently Wrong: A Multi-Model Study of Linear Representations of Synthetic Deception
von: Zolfaghari, Vahideh
Veröffentlicht: (2026)
von: Zolfaghari, Vahideh
Veröffentlicht: (2026)
Unified Class and Domain Incremental Learning with Mixture of Experts for Indoor Localization
von: Singampalli, Akhil, et al.
Veröffentlicht: (2025)
von: Singampalli, Akhil, et al.
Veröffentlicht: (2025)
All or None: Identifiable Linear Properties of Next-token Predictors in Language Modeling
von: Marconato, Emanuele, et al.
Veröffentlicht: (2024)
von: Marconato, Emanuele, et al.
Veröffentlicht: (2024)
On the Separability of Information in Diffusion Models
von: Premkumar, Akhil
Veröffentlicht: (2025)
von: Premkumar, Akhil
Veröffentlicht: (2025)
Easy Problems That LLMs Get Wrong
von: Williams, Sean, et al.
Veröffentlicht: (2024)
von: Williams, Sean, et al.
Veröffentlicht: (2024)
Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?
von: Wen, Xueru, et al.
Veröffentlicht: (2024)
von: Wen, Xueru, et al.
Veröffentlicht: (2024)
All Leaks Count, Some Count More: Interpretable Temporal Contamination Detection and Mitigation in LLM Backtesting
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2026)
When Learning Rates Go Wrong: Early Structural Signals in PPO Actor-Critic
von: Fernández-Hernández, Alberto, et al.
Veröffentlicht: (2026)
von: Fernández-Hernández, Alberto, et al.
Veröffentlicht: (2026)
syftr: Pareto-Optimal Generative AI
von: Conway, Alexander, et al.
Veröffentlicht: (2025)
von: Conway, Alexander, et al.
Veröffentlicht: (2025)
DAILOC: Domain-Incremental Learning for Indoor Localization using Smartphones
von: Singampalli, Akhil, et al.
Veröffentlicht: (2025)
von: Singampalli, Akhil, et al.
Veröffentlicht: (2025)
ACPO: A Policy Optimization Algorithm for Average MDPs with Constraints
von: Agnihotri, Akhil, et al.
Veröffentlicht: (2023)
von: Agnihotri, Akhil, et al.
Veröffentlicht: (2023)
AI Alignment with Changing and Influenceable Reward Functions
von: Carroll, Micah, et al.
Veröffentlicht: (2024)
von: Carroll, Micah, et al.
Veröffentlicht: (2024)
Capabilities Ain't All You Need: Measuring Propensities in AI
von: Romero-Alvarado, Daniel, et al.
Veröffentlicht: (2026)
von: Romero-Alvarado, Daniel, et al.
Veröffentlicht: (2026)
Weak-to-Strong Elicitation via Mismatched Wrong Drafts
von: Deng, Wei
Veröffentlicht: (2026)
von: Deng, Wei
Veröffentlicht: (2026)
Perplexity Cannot Always Tell Right from Wrong
von: Veličković, Petar, et al.
Veröffentlicht: (2026)
von: Veličković, Petar, et al.
Veröffentlicht: (2026)
Loss-to-Loss Prediction: Scaling Laws for All Datasets
von: Brandfonbrener, David, et al.
Veröffentlicht: (2024)
von: Brandfonbrener, David, et al.
Veröffentlicht: (2024)
Uncertainty Quantification in SVM prediction
von: Anand, Pritam
Veröffentlicht: (2025)
von: Anand, Pritam
Veröffentlicht: (2025)
On The Planning Abilities of OpenAI's o1 Models: Feasibility, Optimality, and Generalizability
von: Wang, Kevin, et al.
Veröffentlicht: (2024)
von: Wang, Kevin, et al.
Veröffentlicht: (2024)
Uncertainty Quantification in Probabilistic Machine Learning Models: Theory, Methods, and Insights
von: Ajirak, Marzieh, et al.
Veröffentlicht: (2025)
von: Ajirak, Marzieh, et al.
Veröffentlicht: (2025)
Per-Domain Generalizing Policies: On Learning Efficient and Robust Q-Value Functions (Extended Version with Technical Appendix)
von: Müller, Nicola J., et al.
Veröffentlicht: (2026)
von: Müller, Nicola J., et al.
Veröffentlicht: (2026)
Enactor: From Traffic Simulators to Surrogate World Models
von: Ranjan, Yash, et al.
Veröffentlicht: (2026)
von: Ranjan, Yash, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Closing the Sim2Real Performance Gap in RL
von: Anand, Akhil S, et al.
Veröffentlicht: (2025) -
Economic Model Predictive Control as a Solution to Markov Decision Processes
von: Reinhardt, Dirk, et al.
Veröffentlicht: (2024) -
Data-Driven Predictive Control and MPC: Do we achieve optimality?
von: Anand, Akhil S, et al.
Veröffentlicht: (2024) -
Synthesis of Model Predictive Control and Reinforcement Learning: Survey and Classification
von: Reiter, Rudolf, et al.
Veröffentlicht: (2025) -
Direct transfer of optimized controllers to similar systems using dimensionless MPC
von: Hromatko, Josip Kir, et al.
Veröffentlicht: (2025)