Information maximization for a broad variety of multi-armed bandit games
Fuente:
arXiv
Guardado en:
| Autores principales: | Barbier-Chebbah, Alex, Vestergaard, Christian L., Masson, Jean-Baptiste |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Approximate information maximization for bandit games
por: Barbier-Chebbah, Alex, et al.
Publicado: (2023)
por: Barbier-Chebbah, Alex, et al.
Publicado: (2023)
Long-term memory induced correction to Arrhenius law
por: Barbier-Chebbah, A., et al.
Publicado: (2024)
por: Barbier-Chebbah, A., et al.
Publicado: (2024)
Compression-based inference of network motif sets
por: Bénichou, Alexis, et al.
Publicado: (2023)
por: Bénichou, Alexis, et al.
Publicado: (2023)
Flips Reveal the Universal Impact of Memory on Random Explorations
por: Brémont, Julien, et al.
Publicado: (2025)
por: Brémont, Julien, et al.
Publicado: (2025)
Minimax-optimal trust-aware multi-armed bandits
por: Cai, Changxiao, et al.
Publicado: (2024)
por: Cai, Changxiao, et al.
Publicado: (2024)
Statistical physics of deep learning: Optimal learning of a multi-layer perceptron near interpolation
por: Barbier, Jean, et al.
Publicado: (2025)
por: Barbier, Jean, et al.
Publicado: (2025)
Sharp feature-learning transitions and Bayes-optimal neural scaling laws in extensive-width networks
por: Nguyen, Minh-Toan, et al.
Publicado: (2026)
por: Nguyen, Minh-Toan, et al.
Publicado: (2026)
Machine learning for cerebral blood vessels' malformations
por: Topal, Irem, et al.
Publicado: (2024)
por: Topal, Irem, et al.
Publicado: (2024)
UCB algorithms for multi-armed bandits: Precise regret and adaptive inference
por: Han, Qiyang, et al.
Publicado: (2024)
por: Han, Qiyang, et al.
Publicado: (2024)
The committee machine: Computational to statistical gaps in learning a two-layers neural network
por: Aubin, Benjamin, et al.
Publicado: (2018)
por: Aubin, Benjamin, et al.
Publicado: (2018)
Statistical mechanics of extensive-width Bayesian neural networks near interpolation
por: Barbier, Jean, et al.
Publicado: (2025)
por: Barbier, Jean, et al.
Publicado: (2025)
Optimal generalisation and learning transition in extensive-width shallow neural networks near interpolation
por: Barbier, Jean, et al.
Publicado: (2025)
por: Barbier, Jean, et al.
Publicado: (2025)
Landscape Complexity for the Empirical Risk of Generalized Linear Models: Discrimination between Structured Data
por: Tsironis, Theodoros G., et al.
Publicado: (2025)
por: Tsironis, Theodoros G., et al.
Publicado: (2025)
Learned Mappings for Targeted Free Energy Perturbation between Peptide Conformations
por: Willow, Soohaeng Yoo, et al.
Publicado: (2023)
por: Willow, Soohaeng Yoo, et al.
Publicado: (2023)
Accurate generation of stochastic dynamics based on multi-model Generative Adversarial Networks
por: Lanzoni, Daniele, et al.
Publicado: (2023)
por: Lanzoni, Daniele, et al.
Publicado: (2023)
Learning the Intrinsic Dimensionality of Fermi-Pasta-Ulam-Tsingou Trajectories: A Nonlinear Approach using a Deep Autoencoder Model
por: Marchetti, Gionni
Publicado: (2026)
por: Marchetti, Gionni
Publicado: (2026)
Information-Geometric Decomposition of Generalization Error in Unsupervised Learning
por: Kim, Gilhan
Publicado: (2026)
por: Kim, Gilhan
Publicado: (2026)
Statistics of Min-max Normalized Eigenvalues in Random Matrices
por: Nakada, Hyakka, et al.
Publicado: (2025)
por: Nakada, Hyakka, et al.
Publicado: (2025)
Laws of thermodynamics for exponential families
por: Balsubramani, Akshay
Publicado: (2025)
por: Balsubramani, Akshay
Publicado: (2025)
Susceptibilities and Patterning: A Primer on Linear Response in Bayesian Learning
por: Elliott, Chris, et al.
Publicado: (2026)
por: Elliott, Chris, et al.
Publicado: (2026)
Review and Prospect of Algebraic Research in Equivalent Framework between Statistical Mechanics and Machine Learning Theory
por: Watanabe, Sumio
Publicado: (2024)
por: Watanabe, Sumio
Publicado: (2024)
Scalable Boltzmann Generators for equilibrium sampling of large-scale materials
por: Schebek, Maximilian, et al.
Publicado: (2025)
por: Schebek, Maximilian, et al.
Publicado: (2025)
Emergence of Nonequilibrium Latent Cycles in Unsupervised Generative Modeling
por: Baiesi, Marco, et al.
Publicado: (2025)
por: Baiesi, Marco, et al.
Publicado: (2025)
Machine Learning H-theorem
por: Lier, Ruben
Publicado: (2025)
por: Lier, Ruben
Publicado: (2025)
Dynamical symmetries in the fluctuation-driven regime: an application of Noether's theorem to noisy dynamical systems
por: Vastola, John J.
Publicado: (2025)
por: Vastola, John J.
Publicado: (2025)
Optimised Feature Subset Selection via Simulated Annealing
por: Martínez-García, Fernando, et al.
Publicado: (2025)
por: Martínez-García, Fernando, et al.
Publicado: (2025)
Plastic tensor networks for interpretable generative modeling
por: Akamatsu, Katsuya O., et al.
Publicado: (2025)
por: Akamatsu, Katsuya O., et al.
Publicado: (2025)
Building causation links in stochastic nonlinear systems from data
por: Chibbaro, Sergio, et al.
Publicado: (2025)
por: Chibbaro, Sergio, et al.
Publicado: (2025)
SETOL: A Semi-Empirical Theory of (Deep) Learning
por: Martin, Charles H, et al.
Publicado: (2025)
por: Martin, Charles H, et al.
Publicado: (2025)
Efficient Identification of Critical Transitions via Flow Matching: A Scalable Generative Approach for Many-Body Systems
por: Lee, Qian-Rui, et al.
Publicado: (2025)
por: Lee, Qian-Rui, et al.
Publicado: (2025)
Training thermodynamic computers by gradient descent
por: Whitelam, Stephen
Publicado: (2025)
por: Whitelam, Stephen
Publicado: (2025)
Identifying Ising and percolation phase transitions based on KAN method
por: Xu, Dian, et al.
Publicado: (2025)
por: Xu, Dian, et al.
Publicado: (2025)
High-entropy Advantage in Neural Networks' Generalizability
por: Yang, Entao, et al.
Publicado: (2025)
por: Yang, Entao, et al.
Publicado: (2025)
Optimal Computation from Fluctuation Responses
por: Lyu, Jinghao, et al.
Publicado: (2025)
por: Lyu, Jinghao, et al.
Publicado: (2025)
Thermodynamic Performance Limits for Score-Based Diffusion Models
por: Kodama, Nathan X., et al.
Publicado: (2025)
por: Kodama, Nathan X., et al.
Publicado: (2025)
Improving FMQA via Initial Training Data Design Considering Marginal Bit Coverage in One-Hot Encoding
por: Hayashi, Taiga, et al.
Publicado: (2026)
por: Hayashi, Taiga, et al.
Publicado: (2026)
Combining Reinforcement Learning and Tensor Networks, with an Application to Dynamical Large Deviations
por: Gillman, Edward, et al.
Publicado: (2022)
por: Gillman, Edward, et al.
Publicado: (2022)
Posterior Collapse as Automatic Spectral Pruning
por: Hirn, Johannes
Publicado: (2026)
por: Hirn, Johannes
Publicado: (2026)
The impact of memory on learning sequence-to-sequence tasks
por: Seif, Alireza, et al.
Publicado: (2022)
por: Seif, Alireza, et al.
Publicado: (2022)
Discovering Causal Structure with Reproducing-Kernel Hilbert Space $ε$-Machines
por: Brodu, Nicolas, et al.
Publicado: (2020)
por: Brodu, Nicolas, et al.
Publicado: (2020)
Ejemplares similares
-
Approximate information maximization for bandit games
por: Barbier-Chebbah, Alex, et al.
Publicado: (2023) -
Long-term memory induced correction to Arrhenius law
por: Barbier-Chebbah, A., et al.
Publicado: (2024) -
Compression-based inference of network motif sets
por: Bénichou, Alexis, et al.
Publicado: (2023) -
Flips Reveal the Universal Impact of Memory on Random Explorations
por: Brémont, Julien, et al.
Publicado: (2025) -
Minimax-optimal trust-aware multi-armed bandits
por: Cai, Changxiao, et al.
Publicado: (2024)