Maximum diffusion reinforcement learning
Fuente:
arXiv
Saved in:
| Main Authors: | Berrueta, Thomas A., Pinosky, Allison, Murphey, Todd D. |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Flow Matching Ergodic Coverage
by: Sun, Max Muchen, et al.
Published: (2025)
by: Sun, Max Muchen, et al.
Published: (2025)
Embodied Active Learning of Generative Sensor-Object Models
by: Pinosky, Allison, et al.
Published: (2024)
by: Pinosky, Allison, et al.
Published: (2024)
Cross-fluctuation phase transitions reveal sampling dynamics in diffusion models
by: Ramachandran, Sai Niranjan, et al.
Published: (2025)
by: Ramachandran, Sai Niranjan, et al.
Published: (2025)
Grokking as a Falsifiable Finite-Size Transition
by: Bi, Yuda, et al.
Published: (2026)
by: Bi, Yuda, et al.
Published: (2026)
Entropy, concentration, and learning: a statistical mechanics primer
by: Balsubramani, Akshay
Published: (2024)
by: Balsubramani, Akshay
Published: (2024)
Trees to Flows and Back: Unifying Decision Trees and Diffusion Models
by: Ramachandran, Sai Niranjan, et al.
Published: (2026)
by: Ramachandran, Sai Niranjan, et al.
Published: (2026)
Symmetry and Generalisation in Neural Approximations of Renormalisation Transformations
by: Ashworth, Cassidy, et al.
Published: (2025)
by: Ashworth, Cassidy, et al.
Published: (2025)
Spontaneous symmetry breaking and Goldstone modes for deep information propagation
by: Iqbal, Nabil, et al.
Published: (2026)
by: Iqbal, Nabil, et al.
Published: (2026)
On Distinguishing Capability Elicitation from Capability Creation in Post-Training: A Free-Energy Perspective
by: Li, Yuhao, et al.
Published: (2026)
by: Li, Yuhao, et al.
Published: (2026)
Thermodynamic Irreversibility of Training Algorithms
by: Ziyin, Liu, et al.
Published: (2026)
by: Ziyin, Liu, et al.
Published: (2026)
Reinforcement Learning for Microcanonical Graph Ensemble with Assortativity Constraints
by: Choi, Hoyun, et al.
Published: (2026)
by: Choi, Hoyun, et al.
Published: (2026)
Stochastic Resetting Mitigates Latent Gradient Bias of SGD from Label Noise
by: Bae, Youngkyoung, et al.
Published: (2024)
by: Bae, Youngkyoung, et al.
Published: (2024)
Tensor tree learns hidden relational structures in data to construct generative models
by: Harada, Kenji, et al.
Published: (2024)
by: Harada, Kenji, et al.
Published: (2024)
Intuition emerges in Maximum Caliber models at criticality
by: Arola-Fernández, Lluís
Published: (2025)
by: Arola-Fernández, Lluís
Published: (2025)
Privacy-preserving machine learning with tensor networks
by: Pozas-Kerstjens, Alejandro, et al.
Published: (2022)
by: Pozas-Kerstjens, Alejandro, et al.
Published: (2022)
Machine learning and optimization-based approaches to duality in statistical physics
by: Ferrari, Andrea E. V., et al.
Published: (2024)
by: Ferrari, Andrea E. V., et al.
Published: (2024)
Sampling Decisions
by: Chertkov, Michael, et al.
Published: (2025)
by: Chertkov, Michael, et al.
Published: (2025)
Phase Transitions in the Output Distribution of Large Language Models
by: Arnold, Julian, et al.
Published: (2024)
by: Arnold, Julian, et al.
Published: (2024)
Mixing Artificial and Natural Intelligence: From Statistical Mechanics to AI and Back to Turbulence
by: Chertkov, Michael
Published: (2024)
by: Chertkov, Michael
Published: (2024)
Temporal Memory for Resource-Constrained Agents: Continual Learning via Stochastic Compress-Add-Smooth
by: Chertkov, Michael
Published: (2026)
by: Chertkov, Michael
Published: (2026)
Adaptive Path Integral Diffusion: AdaPID
by: Chertkov, Michael, et al.
Published: (2025)
by: Chertkov, Michael, et al.
Published: (2025)
Scalable Discrete Diffusion Samplers: Combinatorial Optimization and Statistical Physics
by: Sanokowski, Sebastian, et al.
Published: (2025)
by: Sanokowski, Sebastian, et al.
Published: (2025)
IsingFormer: Augmenting Parallel Tempering With Learned Proposals
by: Bunaiyan, Saleh, et al.
Published: (2025)
by: Bunaiyan, Saleh, et al.
Published: (2025)
Relaxation-assisted reverse annealing on nonnegative/binary matrix factorization
by: Haba, Renichiro, et al.
Published: (2025)
by: Haba, Renichiro, et al.
Published: (2025)
On the Separability of Information in Diffusion Models
by: Premkumar, Akhil
Published: (2025)
by: Premkumar, Akhil
Published: (2025)
Generative Stochastic Optimal Transport: Guided Harmonic Path-Integral Diffusion
by: Chertkov, Michael
Published: (2025)
by: Chertkov, Michael
Published: (2025)
Controlling dynamics of stochastic systems with deep reinforcement learning
by: Mukhamadiarov, Ruslan
Published: (2025)
by: Mukhamadiarov, Ruslan
Published: (2025)
The critical slowing down in diffusion models
by: Del Bono, Luca Maria, et al.
Published: (2026)
by: Del Bono, Luca Maria, et al.
Published: (2026)
Autonomous navigation of catheters and guidewires in mechanical thrombectomy using inverse reinforcement learning
by: Robertshaw, Harry, et al.
Published: (2024)
by: Robertshaw, Harry, et al.
Published: (2024)
Simulation-based reinforcement learning for real-world autonomous driving
by: Osiński, Błażej, et al.
Published: (2019)
by: Osiński, Błażej, et al.
Published: (2019)
Computing critical exponents in 3D Ising model via pattern recognition/deep learning approach
by: Burt, Timothy A.
Published: (2024)
by: Burt, Timothy A.
Published: (2024)
Leveraging LLMs for reward function design in reinforcement learning control tasks
by: Cardenoso, Franklin, et al.
Published: (2025)
by: Cardenoso, Franklin, et al.
Published: (2025)
Effectiveness of probabilistic contact tracing in epidemic containment: the role of super-spreaders and transmission path reconstruction
by: Muntoni, A. P., et al.
Published: (2023)
by: Muntoni, A. P., et al.
Published: (2023)
Unsupervised learning of anomalous diffusion data
by: Muñoz-Gil, Gorka, et al.
Published: (2021)
by: Muñoz-Gil, Gorka, et al.
Published: (2021)
Using reinforcement learning to probe the role of feedback in skill acquisition
by: Terpin, Antonio, et al.
Published: (2025)
by: Terpin, Antonio, et al.
Published: (2025)
Scilab-RL: A software framework for efficient reinforcement learning and cognitive modeling research
by: Dohmen, Jan, et al.
Published: (2024)
by: Dohmen, Jan, et al.
Published: (2024)
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
by: Sukhija, Bhavya, et al.
Published: (2024)
by: Sukhija, Bhavya, et al.
Published: (2024)
Physics-informed offline reinforcement learning eliminates catastrophic fuel waste in maritime routing
by: Bora, Aniruddha, et al.
Published: (2026)
by: Bora, Aniruddha, et al.
Published: (2026)
Composing diffusion priors with explicit physical context via generative Gibbs sampling
by: Wang, Weizhou, et al.
Published: (2026)
by: Wang, Weizhou, et al.
Published: (2026)
The impact of memory on learning sequence-to-sequence tasks
by: Seif, Alireza, et al.
Published: (2022)
by: Seif, Alireza, et al.
Published: (2022)
Similar Items
-
Flow Matching Ergodic Coverage
by: Sun, Max Muchen, et al.
Published: (2025) -
Embodied Active Learning of Generative Sensor-Object Models
by: Pinosky, Allison, et al.
Published: (2024) -
Cross-fluctuation phase transitions reveal sampling dynamics in diffusion models
by: Ramachandran, Sai Niranjan, et al.
Published: (2025) -
Grokking as a Falsifiable Finite-Size Transition
by: Bi, Yuda, et al.
Published: (2026) -
Entropy, concentration, and learning: a statistical mechanics primer
by: Balsubramani, Akshay
Published: (2024)