Learning When to Stop: Adaptive Latent Reasoning via Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ning, Alex, Kuo, Yen-Ling, Gomes, Gabe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Visualizing LLM Latent Space Geometry Through Dimensionality Reduction
von: Ning, Alex, et al.
Veröffentlicht: (2025)
von: Ning, Alex, et al.
Veröffentlicht: (2025)
Learning When to Switch: Adaptive Policy Selection via Reinforcement Learning
von: Tava, Chris
Veröffentlicht: (2025)
von: Tava, Chris
Veröffentlicht: (2025)
The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning
von: Liu, Jiashun, et al.
Veröffentlicht: (2025)
von: Liu, Jiashun, et al.
Veröffentlicht: (2025)
Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift
von: Goel, Surbhi, et al.
Veröffentlicht: (2026)
von: Goel, Surbhi, et al.
Veröffentlicht: (2026)
Lessons Learned: Reproducibility, Replicability, and When to Stop
von: Gomez, Milton S., et al.
Veröffentlicht: (2024)
von: Gomez, Milton S., et al.
Veröffentlicht: (2024)
Learning to Reason Efficiently with Discounted Reinforcement Learning
von: Ayoub, Alex, et al.
Veröffentlicht: (2025)
von: Ayoub, Alex, et al.
Veröffentlicht: (2025)
A Theory of Generalization in Deep Learning
von: Litman, Elon, et al.
Veröffentlicht: (2026)
von: Litman, Elon, et al.
Veröffentlicht: (2026)
When to Stop Federated Learning: Zero-Shot Generation of Synthetic Validation Data with Generative AI for Early Stopping
von: Lee, Youngjoon, et al.
Veröffentlicht: (2025)
von: Lee, Youngjoon, et al.
Veröffentlicht: (2025)
Prolonging Tool Life: Learning Skillful Use of General-purpose Tools through Lifespan-guided Reinforcement Learning
von: Wu, Po-Yen, et al.
Veröffentlicht: (2025)
von: Wu, Po-Yen, et al.
Veröffentlicht: (2025)
Adaptive Active Learning for Regression via Reinforcement Learning
von: Nguyen, Simon D., et al.
Veröffentlicht: (2026)
von: Nguyen, Simon D., et al.
Veröffentlicht: (2026)
Self-Reinforced Graph Contrastive Learning
von: Hsieh, Chou-Ying, et al.
Veröffentlicht: (2025)
von: Hsieh, Chou-Ying, et al.
Veröffentlicht: (2025)
Optimal Stopping in Latent Diffusion Models
von: Wu, Yu-Han, et al.
Veröffentlicht: (2025)
von: Wu, Yu-Han, et al.
Veröffentlicht: (2025)
Act Only When It Pays: Efficient Reinforcement Learning for LLM Reasoning via Selective Rollouts
von: Zheng, Haizhong, et al.
Veröffentlicht: (2025)
von: Zheng, Haizhong, et al.
Veröffentlicht: (2025)
Latent-Space Contrastive Reinforcement Learning for Stable and Efficient LLM Reasoning
von: Shan, Lianlei, et al.
Veröffentlicht: (2026)
von: Shan, Lianlei, et al.
Veröffentlicht: (2026)
Learning to Stop: Deep Learning for Mean Field Optimal Stopping
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2024)
von: Magnino, Lorenzo, et al.
Veröffentlicht: (2024)
When Actions Teach You to Think: Reasoning-Action Synergy via Reinforcement Learning in Conversational Agents
von: Rawat, Mrinal, et al.
Veröffentlicht: (2025)
von: Rawat, Mrinal, et al.
Veröffentlicht: (2025)
Teaching Large Language Models to Reason with Reinforcement Learning
von: Havrilla, Alex, et al.
Veröffentlicht: (2024)
von: Havrilla, Alex, et al.
Veröffentlicht: (2024)
Quantum Reinforcement Learning by Adaptive Non-local Observables
von: Lin, Hsin-Yi, et al.
Veröffentlicht: (2025)
von: Lin, Hsin-Yi, et al.
Veröffentlicht: (2025)
Random Latent Exploration for Deep Reinforcement Learning
von: Mahankali, Srinath, et al.
Veröffentlicht: (2024)
von: Mahankali, Srinath, et al.
Veröffentlicht: (2024)
Latent Poincaré Shaping for Agentic Reinforcement Learning
von: Xia, Hanchen, et al.
Veröffentlicht: (2026)
von: Xia, Hanchen, et al.
Veröffentlicht: (2026)
Learning When to Act: Communication-Efficient Reinforcement Learning via Run-Time Assurance
von: Haroon, Adam, et al.
Veröffentlicht: (2026)
von: Haroon, Adam, et al.
Veröffentlicht: (2026)
AdaptThink: Reasoning Models Can Learn When to Think
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
Multi-Objective Adaptive Rate Limiting in Microservices Using Deep Reinforcement Learning
von: Lyu, Ning, et al.
Veröffentlicht: (2025)
von: Lyu, Ning, et al.
Veröffentlicht: (2025)
Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models
von: Nath, Vaskar, et al.
Veröffentlicht: (2025)
von: Nath, Vaskar, et al.
Veröffentlicht: (2025)
Early Stopping Tabular In-Context Learning
von: Küken, Jaris, et al.
Veröffentlicht: (2025)
von: Küken, Jaris, et al.
Veröffentlicht: (2025)
Vision Transformers that Never Stop Learning
von: Sun, Caihao, et al.
Veröffentlicht: (2026)
von: Sun, Caihao, et al.
Veröffentlicht: (2026)
LaDi-RL: Latent Diffusion Reasoning Prevents Entropy Collapse in Reinforcement Learning
von: Kang, Haoqiang, et al.
Veröffentlicht: (2026)
von: Kang, Haoqiang, et al.
Veröffentlicht: (2026)
AdaMemento: Adaptive Memory-Assisted Policy Optimization for Reinforcement Learning
von: Yan, Renye, et al.
Veröffentlicht: (2024)
von: Yan, Renye, et al.
Veröffentlicht: (2024)
The Hidden Influence of Latent Feature Magnitude When Learning with Imbalanced Data
von: Dablain, Damien A., et al.
Veröffentlicht: (2024)
von: Dablain, Damien A., et al.
Veröffentlicht: (2024)
Safe Reinforcement Learning for Constrained Markov Decision Processes with Stochastic Stopping Time
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
von: Mazumdar, Abhijit, et al.
Veröffentlicht: (2024)
When Does Neuroevolution Outcompete Reinforcement Learning in Transfer Learning Tasks?
von: Nisioti, Eleni, et al.
Veröffentlicht: (2025)
von: Nisioti, Eleni, et al.
Veröffentlicht: (2025)
Rich-Observation Reinforcement Learning with Continuous Latent Dynamics
von: Song, Yuda, et al.
Veröffentlicht: (2024)
von: Song, Yuda, et al.
Veröffentlicht: (2024)
Adaptive Robust Learning using Latent Bernoulli Variables
von: Karakulev, Aleksandr, et al.
Veröffentlicht: (2023)
von: Karakulev, Aleksandr, et al.
Veröffentlicht: (2023)
Quantum Hierarchical Reinforcement Learning via Variational Quantum Circuits
von: Lee, Yu-Ting, et al.
Veröffentlicht: (2026)
von: Lee, Yu-Ting, et al.
Veröffentlicht: (2026)
A Reinforcement Learning Inspired Latent Yield Based Adaptive Algorithm Switching Mechanism
von: Nair, Jayprakash S., et al.
Veröffentlicht: (2026)
von: Nair, Jayprakash S., et al.
Veröffentlicht: (2026)
Reinforcement Learning for Adaptive MCMC
von: Wang, Congye, et al.
Veröffentlicht: (2024)
von: Wang, Congye, et al.
Veröffentlicht: (2024)
Multi-Path Collaborative Reasoning via Reinforcement Learning
von: Lv, Jindi, et al.
Veröffentlicht: (2025)
von: Lv, Jindi, et al.
Veröffentlicht: (2025)
Efficient Reinforcement Finetuning via Adaptive Curriculum Learning
von: Shi, Taiwei, et al.
Veröffentlicht: (2025)
von: Shi, Taiwei, et al.
Veröffentlicht: (2025)
Advancing Molecular Machine Learning Representations with Stereoelectronics-Infused Molecular Graphs
von: Boiko, Daniil A., et al.
Veröffentlicht: (2024)
von: Boiko, Daniil A., et al.
Veröffentlicht: (2024)
Autonomous Adaptive Solver Selection for Chemistry Integration via Reinforcement Learning
von: Ikponmwoba, Eloghosa, et al.
Veröffentlicht: (2026)
von: Ikponmwoba, Eloghosa, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Visualizing LLM Latent Space Geometry Through Dimensionality Reduction
von: Ning, Alex, et al.
Veröffentlicht: (2025) -
Learning When to Switch: Adaptive Policy Selection via Reinforcement Learning
von: Tava, Chris
Veröffentlicht: (2025) -
The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning
von: Liu, Jiashun, et al.
Veröffentlicht: (2025) -
Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift
von: Goel, Surbhi, et al.
Veröffentlicht: (2026) -
Lessons Learned: Reproducibility, Replicability, and When to Stop
von: Gomez, Milton S., et al.
Veröffentlicht: (2024)