Towards a Law of Iterated Expectations for Heuristic Estimators
Fuente:
arXiv
Guardado en:
| Autores principales: | Christiano, Paul, Hilton, Jacob, Lincoln, Andrea, Neyman, Eric, Xu, Mark |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Backdoor defense, learnability and obfuscation
por: Christiano, Paul, et al.
Publicado: (2024)
por: Christiano, Paul, et al.
Publicado: (2024)
Estimating the Probabilities of Rare Outputs in Language Models
por: Wu, Gabriel, et al.
Publicado: (2024)
por: Wu, Gabriel, et al.
Publicado: (2024)
Circuits, Features, and Heuristics in Molecular Transformers
por: Varadi, Kristof, et al.
Publicado: (2025)
por: Varadi, Kristof, et al.
Publicado: (2025)
A Hitchhiker's Guide to Scaling Law Estimation
por: Choshen, Leshem, et al.
Publicado: (2024)
por: Choshen, Leshem, et al.
Publicado: (2024)
Towards Learning Foundation Models for Heuristic Functions to Solve Pathfinding Problems
por: Khandelwal, Vedant, et al.
Publicado: (2024)
por: Khandelwal, Vedant, et al.
Publicado: (2024)
Expectation Alignment: Handling Reward Misspecification in the Presence of Expectation Mismatch
por: Mechergui, Malek, et al.
Publicado: (2024)
por: Mechergui, Malek, et al.
Publicado: (2024)
Algorithmic Bayesian Epistemology
por: Neyman, Eric
Publicado: (2024)
por: Neyman, Eric
Publicado: (2024)
Learning Social Heuristics for Human-Aware Path Planning
por: Eirale, Andrea, et al.
Publicado: (2025)
por: Eirale, Andrea, et al.
Publicado: (2025)
Towards Neural Scaling Laws on Graphs
por: Liu, Jingzhe, et al.
Publicado: (2024)
por: Liu, Jingzhe, et al.
Publicado: (2024)
Expectation Maximization Pseudo Labels
por: Xu, Moucheng, et al.
Publicado: (2023)
por: Xu, Moucheng, et al.
Publicado: (2023)
Global Sensitivity Analysis for Engineering Design Based on Individual Conditional Expectations
por: Palar, Pramudita Satria, et al.
Publicado: (2025)
por: Palar, Pramudita Satria, et al.
Publicado: (2025)
Wukong: Towards a Scaling Law for Large-Scale Recommendation
por: Zhang, Buyun, et al.
Publicado: (2024)
por: Zhang, Buyun, et al.
Publicado: (2024)
Reinforcement Learning with Conditional Expectation Reward
por: Xiao, Changyi, et al.
Publicado: (2026)
por: Xiao, Changyi, et al.
Publicado: (2026)
Neural Expectation Operators
por: Qi, Qian
Publicado: (2025)
por: Qi, Qian
Publicado: (2025)
Bayesian Deep Learning Via Expectation Maximization and Turbo Deep Approximate Message Passing
por: Xu, Wei, et al.
Publicado: (2024)
por: Xu, Wei, et al.
Publicado: (2024)
Towards a Comprehensive Scaling Law of Mixture-of-Experts
por: Zhao, Guoliang, et al.
Publicado: (2025)
por: Zhao, Guoliang, et al.
Publicado: (2025)
Towards Neural Scaling Laws for Time Series Foundation Models
por: Yao, Qingren, et al.
Publicado: (2024)
por: Yao, Qingren, et al.
Publicado: (2024)
Going Beyond Heuristics by Imposing Policy Improvement as a Constraint
por: Lee, Chi-Chang, et al.
Publicado: (2025)
por: Lee, Chi-Chang, et al.
Publicado: (2025)
Demystifying the Recency Heuristic in Temporal-Difference Learning
por: Daley, Brett, et al.
Publicado: (2024)
por: Daley, Brett, et al.
Publicado: (2024)
Learning Admissible Heuristics for A*: Theory and Practice
por: Futuhi, Ehsan, et al.
Publicado: (2025)
por: Futuhi, Ehsan, et al.
Publicado: (2025)
GEPO: Group Expectation Policy Optimization for Stable Heterogeneous Reinforcement Learning
por: Zhang, Han, et al.
Publicado: (2025)
por: Zhang, Han, et al.
Publicado: (2025)
Simulation-Free Differential Dynamics through Neural Conservation Laws
por: Hua, Mengjian, et al.
Publicado: (2025)
por: Hua, Mengjian, et al.
Publicado: (2025)
Towards Embodiment Scaling Laws in Robot Locomotion
por: Ai, Bo, et al.
Publicado: (2025)
por: Ai, Bo, et al.
Publicado: (2025)
Toward a Metrology for Artificial Intelligence: Hidden-Rule Environments and Reinforcement Learning
por: Mathew, Christo, et al.
Publicado: (2025)
por: Mathew, Christo, et al.
Publicado: (2025)
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs
por: Bian, Song, et al.
Publicado: (2025)
por: Bian, Song, et al.
Publicado: (2025)
A Unified Theory of $θ$-Expectations
por: Qi, Qian
Publicado: (2025)
por: Qi, Qian
Publicado: (2025)
PALATE: Peculiar Application of the Law of Total Expectation to Enhance the Evaluation of Deep Generative Models
por: Dziarmaga, Tadeusz, et al.
Publicado: (2025)
por: Dziarmaga, Tadeusz, et al.
Publicado: (2025)
Deep Heuristic Learning for Real-Time Urban Pathfinding
por: El-Ela, Mohamed Hussein Abo, et al.
Publicado: (2024)
por: El-Ela, Mohamed Hussein Abo, et al.
Publicado: (2024)
Enhancing Q-Learning with Large Language Model Heuristics
por: Wu, Xiefeng
Publicado: (2024)
por: Wu, Xiefeng
Publicado: (2024)
Heuristic Transformer: Belief Augmented In-Context Reinforcement Learning
por: Dippel, Oliver, et al.
Publicado: (2025)
por: Dippel, Oliver, et al.
Publicado: (2025)
Generalizable Heuristic Generation Through LLMs with Meta-Optimization
por: Shi, Yiding, et al.
Publicado: (2025)
por: Shi, Yiding, et al.
Publicado: (2025)
An Expectation-Maximization Algorithm for Domain Adaptation in Gaussian Causal Models
por: Javidian, Mohammad Ali
Publicado: (2026)
por: Javidian, Mohammad Ali
Publicado: (2026)
Subliminal Learning: Language models transmit behavioral traits via hidden signals in data
por: Cloud, Alex, et al.
Publicado: (2025)
por: Cloud, Alex, et al.
Publicado: (2025)
Towards Anytime-Valid Statistical Watermarking
por: Huang, Baihe, et al.
Publicado: (2026)
por: Huang, Baihe, et al.
Publicado: (2026)
A Benchmark for Maximum Cut: Towards Standardization of the Evaluation of Learned Heuristics for Combinatorial Optimization
por: Nath, Ankur, et al.
Publicado: (2024)
por: Nath, Ankur, et al.
Publicado: (2024)
Learning a Generic Value-Selection Heuristic Inside a Constraint Programming Solver
por: Marty, Tom, et al.
Publicado: (2023)
por: Marty, Tom, et al.
Publicado: (2023)
Using Scaling Laws for Data Source Utility Estimation in Domain-Specific Pre-Training
por: Ostapenko, Oleksiy, et al.
Publicado: (2025)
por: Ostapenko, Oleksiy, et al.
Publicado: (2025)
EXGnet: a single-lead explainable-AI guided multiresolution network with train-only quantitative features for trustworthy ECG arrhythmia classification
por: Showrav, Tushar Talukder, et al.
Publicado: (2025)
por: Showrav, Tushar Talukder, et al.
Publicado: (2025)
Adaptive Variational Continual Learning via Task-Heuristic Modelling
por: Yang, Fan
Publicado: (2024)
por: Yang, Fan
Publicado: (2024)
Purifying Shampoo: Investigating Shampoo's Heuristics by Decomposing its Preconditioner
por: Eschenhagen, Runa, et al.
Publicado: (2025)
por: Eschenhagen, Runa, et al.
Publicado: (2025)
Ejemplares similares
-
Backdoor defense, learnability and obfuscation
por: Christiano, Paul, et al.
Publicado: (2024) -
Estimating the Probabilities of Rare Outputs in Language Models
por: Wu, Gabriel, et al.
Publicado: (2024) -
Circuits, Features, and Heuristics in Molecular Transformers
por: Varadi, Kristof, et al.
Publicado: (2025) -
A Hitchhiker's Guide to Scaling Law Estimation
por: Choshen, Leshem, et al.
Publicado: (2024) -
Towards Learning Foundation Models for Heuristic Functions to Solve Pathfinding Problems
por: Khandelwal, Vedant, et al.
Publicado: (2024)