Guardado en:
| Autores principales: | Sinha, Kaushik, Tosh, Christopher |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.06175 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
What is causal about causal models and representations?
por: Jørgensen, Frederik Hytting, et al.
Publicado: (2025)
por: Jørgensen, Frederik Hytting, et al.
Publicado: (2025)
Beyond identifiability: Learning causal representations with few environments and finite samples
por: Lee, Inbeom, et al.
Publicado: (2026)
por: Lee, Inbeom, et al.
Publicado: (2026)
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
por: Ji, Kaixuan, et al.
Publicado: (2026)
por: Ji, Kaixuan, et al.
Publicado: (2026)
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
por: Ji, Kaixuan, et al.
Publicado: (2026)
por: Ji, Kaixuan, et al.
Publicado: (2026)
Minimax Optimal Variance-Aware Regret Bounds for Multinomial Logistic MDPs
por: Boudart, Pierre, et al.
Publicado: (2026)
por: Boudart, Pierre, et al.
Publicado: (2026)
Enjoying Non-linearity in Multinomial Logistic Bandits: A Minimax-Optimal Algorithm
por: Boudart, Pierre, et al.
Publicado: (2025)
por: Boudart, Pierre, et al.
Publicado: (2025)
Path Regularization: A Near-Complete and Optimal Nonasymptotic Generalization Theory for Multilayer Neural Networks and Double Descent Phenomenon
por: Yu, Hao
Publicado: (2025)
por: Yu, Hao
Publicado: (2025)
MESSY Estimation: Maximum-Entropy based Stochastic and Symbolic densitY Estimation
por: Tohme, Tony, et al.
Publicado: (2023)
por: Tohme, Tony, et al.
Publicado: (2023)
Statistical Inference for Optimal Transport Maps: Recent Advances and Perspectives
por: Balakrishnan, Sivaraman, et al.
Publicado: (2025)
por: Balakrishnan, Sivaraman, et al.
Publicado: (2025)
A Theory of the Mechanics of Information: Generalization Through Measurement of Uncertainty (Learning is Measuring)
por: Hazard, Christopher J., et al.
Publicado: (2025)
por: Hazard, Christopher J., et al.
Publicado: (2025)
Near-Optimal Learning and Planning in Separated Latent MDPs
por: Chen, Fan, et al.
Publicado: (2024)
por: Chen, Fan, et al.
Publicado: (2024)
Training Dynamics of Multi-Head Softmax Attention for In-Context Learning: Emergence, Convergence, and Optimality
por: Chen, Siyu, et al.
Publicado: (2024)
por: Chen, Siyu, et al.
Publicado: (2024)
On the number of modes of Gaussian kernel density estimators
por: Geshkovski, Borjan, et al.
Publicado: (2024)
por: Geshkovski, Borjan, et al.
Publicado: (2024)
Revisiting Incremental Stochastic Majorization-Minimization Algorithms with Applications to Mixture of Experts
por: Tran, TrungKhang, et al.
Publicado: (2026)
por: Tran, TrungKhang, et al.
Publicado: (2026)
Identifiability of Potentially Degenerate Gaussian Mixture Models With Piecewise Affine Mixing
por: Xu, Danru, et al.
Publicado: (2026)
por: Xu, Danru, et al.
Publicado: (2026)
Generalization Properties of Score-matching Diffusion Models for Intrinsically Low-dimensional Data
por: Chakraborty, Saptarshi, et al.
Publicado: (2026)
por: Chakraborty, Saptarshi, et al.
Publicado: (2026)
A Fine-Grained Understanding of Uniform Convergence for Halfspaces
por: Kontorovich, Aryeh, et al.
Publicado: (2026)
por: Kontorovich, Aryeh, et al.
Publicado: (2026)
Labels or Preferences? Budget-Constrained Learning with Human Judgments over AI-Generated Outputs
por: Dong, Zihan, et al.
Publicado: (2026)
por: Dong, Zihan, et al.
Publicado: (2026)
Finite-Particle Convergence Rates for Conservative and Non-Conservative Drifting Models
por: Balasubramanian, Krishnakumar
Publicado: (2026)
por: Balasubramanian, Krishnakumar
Publicado: (2026)
Low-Dimensional Adaptation of Rectified Flow: A Diffusion and Stochastic Localization Perspective
por: Roy, Saptarshi, et al.
Publicado: (2026)
por: Roy, Saptarshi, et al.
Publicado: (2026)
Ordinary Least Squares is a Special Case of Transformer
por: Tan, Xiaojun, et al.
Publicado: (2026)
por: Tan, Xiaojun, et al.
Publicado: (2026)
A Unified Pair-GRPO Family: From Implicit to Explicit Preference Constraints for Stable and General RL Alignment
por: Yu, Hao
Publicado: (2026)
por: Yu, Hao
Publicado: (2026)
Statistical inference with belief functions: A survey
por: Cuzzolin, Fabio
Publicado: (2026)
por: Cuzzolin, Fabio
Publicado: (2026)
Conformal Policy Control
por: Prinster, Drew, et al.
Publicado: (2026)
por: Prinster, Drew, et al.
Publicado: (2026)
Stein-Rule Shrinkage for Stochastic Gradient Estimation in High Dimensions
por: Arashi, M., et al.
Publicado: (2026)
por: Arashi, M., et al.
Publicado: (2026)
Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees under Contaminated Recursive Training
por: Wang, Kevin, et al.
Publicado: (2026)
por: Wang, Kevin, et al.
Publicado: (2026)
Chemical Reaction Networks Learn Better than Spiking Neural Networks
por: Jaffard, Sophie, et al.
Publicado: (2026)
por: Jaffard, Sophie, et al.
Publicado: (2026)
Total Variation Rates for Riemannian Flow Matching
por: Guan, Yunrui, et al.
Publicado: (2026)
por: Guan, Yunrui, et al.
Publicado: (2026)
LIBRA: Language Model Informed Bandit Recourse Algorithm for Personalized Treatment Planning
por: Cao, Junyu, et al.
Publicado: (2026)
por: Cao, Junyu, et al.
Publicado: (2026)
RACER: Risk-Aware Calibrated Efficient Routing for Large Language Models
por: Hao, Sai, et al.
Publicado: (2026)
por: Hao, Sai, et al.
Publicado: (2026)
A Diffusion Analysis of Policy Gradient for Stochastic Bandits
por: Lattimore, Tor
Publicado: (2026)
por: Lattimore, Tor
Publicado: (2026)
Cost-optimal Sequential Testing via Doubly Robust Q-learning
por: Zhou, Doudou, et al.
Publicado: (2026)
por: Zhou, Doudou, et al.
Publicado: (2026)
A Quantitative Characterization of Forgetting in Post-Training
por: Balasubramanian, Krishnakumar, et al.
Publicado: (2026)
por: Balasubramanian, Krishnakumar, et al.
Publicado: (2026)
Generalization and Scaling Laws for Mixture-of-Experts Transformers
por: Mayaki, Mansour Zoubeirou a
Publicado: (2026)
por: Mayaki, Mansour Zoubeirou a
Publicado: (2026)
Beyond Demand Estimation: Consumer Surplus Evaluation via Cumulative Propensity Weights
por: Bian, Zeyu, et al.
Publicado: (2026)
por: Bian, Zeyu, et al.
Publicado: (2026)
Adaptive auditing of AI systems with anytime-valid guarantees
por: Zhou, Siyu, et al.
Publicado: (2026)
por: Zhou, Siyu, et al.
Publicado: (2026)
Influence functions and regularity tangents for efficient active learning
por: Eaton, Frederik
Publicado: (2024)
por: Eaton, Frederik
Publicado: (2024)
Efficient Knowledge Distillation via Curriculum Extraction
por: Gupta, Shivam, et al.
Publicado: (2025)
por: Gupta, Shivam, et al.
Publicado: (2025)
Risk Analysis and Design Against Adversarial Actions
por: Campi, Marco C., et al.
Publicado: (2025)
por: Campi, Marco C., et al.
Publicado: (2025)
From Spikes to Heavy Tails: Unveiling the Spectral Evolution of Neural Networks
por: Kothapalli, Vignesh, et al.
Publicado: (2024)
por: Kothapalli, Vignesh, et al.
Publicado: (2024)
Ejemplares similares
-
What is causal about causal models and representations?
por: Jørgensen, Frederik Hytting, et al.
Publicado: (2025) -
Beyond identifiability: Learning causal representations with few environments and finite samples
por: Lee, Inbeom, et al.
Publicado: (2026) -
Near-Optimal Regret for KL-Regularized Multi-Armed Bandits
por: Ji, Kaixuan, et al.
Publicado: (2026) -
On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization
por: Ji, Kaixuan, et al.
Publicado: (2026) -
Minimax Optimal Variance-Aware Regret Bounds for Multinomial Logistic MDPs
por: Boudart, Pierre, et al.
Publicado: (2026)