A Computational Theory for Efficient Mini Agent Evaluation with Causal Guarantees
Fuente:
arXiv
Guardado en:
| Autor principal: | Yan, Hedong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On the Provable Performance Guarantee of Efficient Reasoning Models
por: Zeng, Hao, et al.
Publicado: (2025)
por: Zeng, Hao, et al.
Publicado: (2025)
O(d/T) Convergence Theory for Diffusion Probabilistic Models under Minimal Assumptions
por: Li, Gen, et al.
Publicado: (2024)
por: Li, Gen, et al.
Publicado: (2024)
Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees under Contaminated Recursive Training
por: Wang, Kevin, et al.
Publicado: (2026)
por: Wang, Kevin, et al.
Publicado: (2026)
Counterfactual Generative Modeling with Variational Causal Inference
por: Wu, Yulun, et al.
Publicado: (2024)
por: Wu, Yulun, et al.
Publicado: (2024)
Learning Interpretable Concepts: Unifying Causal Representation Learning and Foundation Models
por: Rajendran, Goutham, et al.
Publicado: (2024)
por: Rajendran, Goutham, et al.
Publicado: (2024)
Variational Causal Inference
por: Wu, Yulun, et al.
Publicado: (2022)
por: Wu, Yulun, et al.
Publicado: (2022)
A Theory of the Mechanics of Information: Generalization Through Measurement of Uncertainty (Learning is Measuring)
por: Hazard, Christopher J., et al.
Publicado: (2025)
por: Hazard, Christopher J., et al.
Publicado: (2025)
Diffusion Posterior Sampling is Computationally Intractable
por: Gupta, Shivam, et al.
Publicado: (2024)
por: Gupta, Shivam, et al.
Publicado: (2024)
Foundations of Structural Causal Models with Latent Selection
por: Chen, Leihao, et al.
Publicado: (2024)
por: Chen, Leihao, et al.
Publicado: (2024)
Psychometric Tests for AI Agents and Their Moduli Space
por: Chojecki, Przemyslaw
Publicado: (2025)
por: Chojecki, Przemyslaw
Publicado: (2025)
Scaling Laws in Linear Regression: Compute, Parameters, and Data
por: Lin, Licong, et al.
Publicado: (2024)
por: Lin, Licong, et al.
Publicado: (2024)
Maximizing the Potential of Synthetic Data: Insights from Random Matrix Theory
por: Firdoussi, Aymane El, et al.
Publicado: (2024)
por: Firdoussi, Aymane El, et al.
Publicado: (2024)
Path Regularization: A Near-Complete and Optimal Nonasymptotic Generalization Theory for Multilayer Neural Networks and Double Descent Phenomenon
por: Yu, Hao
Publicado: (2025)
por: Yu, Hao
Publicado: (2025)
Guaranteed Recovery of Unambiguous Clusters
por: Mazooji, Kayvon, et al.
Publicado: (2025)
por: Mazooji, Kayvon, et al.
Publicado: (2025)
Efficient Knowledge Distillation via Curriculum Extraction
por: Gupta, Shivam, et al.
Publicado: (2025)
por: Gupta, Shivam, et al.
Publicado: (2025)
Adapting to Unknown Low-Dimensional Structures in Score-Based Diffusion Models
por: Li, Gen, et al.
Publicado: (2024)
por: Li, Gen, et al.
Publicado: (2024)
RACER: Risk-Aware Calibrated Efficient Routing for Large Language Models
por: Hao, Sai, et al.
Publicado: (2026)
por: Hao, Sai, et al.
Publicado: (2026)
How Particle-System Random Batch Methods Enhance Graph Transformer: Memory Efficiency and Parallel Computing Strategy
por: Liu, Hanwen, et al.
Publicado: (2025)
por: Liu, Hanwen, et al.
Publicado: (2025)
Towards Efficient Online Exploration for Reinforcement Learning with Human Feedback
por: Li, Gen, et al.
Publicado: (2025)
por: Li, Gen, et al.
Publicado: (2025)
Generalization Bounds: Perspectives from Information Theory and PAC-Bayes
por: Hellström, Fredrik, et al.
Publicado: (2023)
por: Hellström, Fredrik, et al.
Publicado: (2023)
U-Nets as Belief Propagation: Efficient Classification, Denoising, and Diffusion in Generative Hierarchical Models
por: Mei, Song
Publicado: (2024)
por: Mei, Song
Publicado: (2024)
Beyond Demand Estimation: Consumer Surplus Evaluation via Cumulative Propensity Weights
por: Bian, Zeyu, et al.
Publicado: (2026)
por: Bian, Zeyu, et al.
Publicado: (2026)
Is a Good Foundation Necessary for Efficient Reinforcement Learning? The Computational Role of the Base Model in Exploration
por: Foster, Dylan J., et al.
Publicado: (2025)
por: Foster, Dylan J., et al.
Publicado: (2025)
Improved Guarantees for Heterogeneous Treatment-Effect Estimation via Matrix Completion
por: Mehrotra, Anay, et al.
Publicado: (2026)
por: Mehrotra, Anay, et al.
Publicado: (2026)
A Score-Based Density Formula, with Applications in Diffusion Generative Models
por: Li, Gen, et al.
Publicado: (2024)
por: Li, Gen, et al.
Publicado: (2024)
Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals
por: Dong, Zihan, et al.
Publicado: (2026)
por: Dong, Zihan, et al.
Publicado: (2026)
Performative Learning Theory
por: Rodemann, Julian, et al.
Publicado: (2026)
por: Rodemann, Julian, et al.
Publicado: (2026)
Foundations of Top-$k$ Decoding For Language Models
por: Noarov, Georgy, et al.
Publicado: (2025)
por: Noarov, Georgy, et al.
Publicado: (2025)
Pessimism in the Face of Confounders: Provably Efficient Offline Reinforcement Learning in Partially Observable Markov Decision Processes
por: Lu, Miao, et al.
Publicado: (2022)
por: Lu, Miao, et al.
Publicado: (2022)
Statistical inference with belief functions: A survey
por: Cuzzolin, Fabio
Publicado: (2026)
por: Cuzzolin, Fabio
Publicado: (2026)
A Quantitative Characterization of Forgetting in Post-Training
por: Balasubramanian, Krishnakumar, et al.
Publicado: (2026)
por: Balasubramanian, Krishnakumar, et al.
Publicado: (2026)
The Geometry of Benchmarks: A New Path Toward AGI
por: Chojecki, Przemyslaw
Publicado: (2025)
por: Chojecki, Przemyslaw
Publicado: (2025)
Deep Ensembles for Epistemic Uncertainty: A Frequentist Perspective
por: Jain, Anchit, et al.
Publicado: (2025)
por: Jain, Anchit, et al.
Publicado: (2025)
A Fine-Grained Understanding of Uniform Convergence for Halfspaces
por: Kontorovich, Aryeh, et al.
Publicado: (2026)
por: Kontorovich, Aryeh, et al.
Publicado: (2026)
A Diffusion Analysis of Policy Gradient for Stochastic Bandits
por: Lattimore, Tor
Publicado: (2026)
por: Lattimore, Tor
Publicado: (2026)
Enjoying Non-linearity in Multinomial Logistic Bandits: A Minimax-Optimal Algorithm
por: Boudart, Pierre, et al.
Publicado: (2025)
por: Boudart, Pierre, et al.
Publicado: (2025)
A note on the impossibility of conditional PAC-efficient reasoning in large language models
por: Zeng, Hao
Publicado: (2025)
por: Zeng, Hao
Publicado: (2025)
A Statistical Analysis of Deep Federated Learning for Intrinsically Low-dimensional Data
por: Chakraborty, Saptarshi, et al.
Publicado: (2024)
por: Chakraborty, Saptarshi, et al.
Publicado: (2024)
Low-Dimensional Adaptation of Rectified Flow: A Diffusion and Stochastic Localization Perspective
por: Roy, Saptarshi, et al.
Publicado: (2026)
por: Roy, Saptarshi, et al.
Publicado: (2026)
A comparative study of conformal prediction methods for valid uncertainty quantification in machine learning
por: Dewolf, Nicolas
Publicado: (2024)
por: Dewolf, Nicolas
Publicado: (2024)
Ejemplares similares
-
On the Provable Performance Guarantee of Efficient Reasoning Models
por: Zeng, Hao, et al.
Publicado: (2025) -
O(d/T) Convergence Theory for Diffusion Probabilistic Models under Minimal Assumptions
por: Li, Gen, et al.
Publicado: (2024) -
Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees under Contaminated Recursive Training
por: Wang, Kevin, et al.
Publicado: (2026) -
Counterfactual Generative Modeling with Variational Causal Inference
por: Wu, Yulun, et al.
Publicado: (2024) -
Learning Interpretable Concepts: Unifying Causal Representation Learning and Foundation Models
por: Rajendran, Goutham, et al.
Publicado: (2024)