PEPS: Quantum-Inspired Reinforcement Learning for Coherent Reasoning Traces in LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Margapuri, Venkat, Kazanjian, Garik, Kosaraju, Naren |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hybrid Safety Verification of Multi-Agent Systems using $ψ$-Weighted CBFs and PAC Guarantees
por: Margapuri, Venkat, et al.
Publicado: (2025)
por: Margapuri, Venkat, et al.
Publicado: (2025)
Diagnosis and Severity Assessment of Ulcerative Colitis using Self Supervised Learning
por: Margapuri, Venkat
Publicado: (2024)
por: Margapuri, Venkat
Publicado: (2024)
Predicting Mortality and Functional Status Scores of Traumatic Brain Injury Patients using Supervised Machine Learning
por: Steinmetz, Lucas, et al.
Publicado: (2024)
por: Steinmetz, Lucas, et al.
Publicado: (2024)
Seed Kernel Counting using Domain Randomization and Object Tracking Neural Networks
por: Margapuri, Venkat, et al.
Publicado: (2023)
por: Margapuri, Venkat, et al.
Publicado: (2023)
Prompt Informed Reinforcement Learning for Visual Coverage Path Planning
por: Margapuri, Venkat
Publicado: (2025)
por: Margapuri, Venkat
Publicado: (2025)
Leaf Angle Estimation using Mask R-CNN and LETR Vision Transformer
por: Margapuri, Venkat, et al.
Publicado: (2024)
por: Margapuri, Venkat, et al.
Publicado: (2024)
TSSR: Two-Stage Swap-Reward-Driven Reinforcement Learning for Character-Level SMILES Generation
por: Levine, Jacob Ede, et al.
Publicado: (2026)
por: Levine, Jacob Ede, et al.
Publicado: (2026)
LLMs as Layout Designers: Enhanced Spatial Reasoning for Content-Aware Layout Generation
por: Li, Sha, et al.
Publicado: (2025)
por: Li, Sha, et al.
Publicado: (2025)
Agentic Reasoning and Tool Integration for LLMs via Reinforcement Learning
por: Singh, Joykirat, et al.
Publicado: (2025)
por: Singh, Joykirat, et al.
Publicado: (2025)
Answering the Wrong Question: Reasoning Trace Inversion for Abstention in LLMs
por: Gourabathina, Abinitha, et al.
Publicado: (2026)
por: Gourabathina, Abinitha, et al.
Publicado: (2026)
Beyond Accuracy: Dissecting Mathematical Reasoning for LLMs Under Reinforcement Learning
por: Wang, Jiayu, et al.
Publicado: (2025)
por: Wang, Jiayu, et al.
Publicado: (2025)
Emergent Hierarchical Reasoning in LLMs through Reinforcement Learning
por: Wang, Haozhe, et al.
Publicado: (2025)
por: Wang, Haozhe, et al.
Publicado: (2025)
BERT Learns (and Teaches) Chemistry
por: Payne, Josh, et al.
Publicado: (2020)
por: Payne, Josh, et al.
Publicado: (2020)
Hidden Thoughts Are Not Secret: Reasoning Trace Exposure in LLMs
por: Lu, Yu-An, et al.
Publicado: (2026)
por: Lu, Yu-An, et al.
Publicado: (2026)
Toward Better EHR Reasoning in LLMs: Reinforcement Learning with Expert Attention Guidance
por: Fang, Yue, et al.
Publicado: (2025)
por: Fang, Yue, et al.
Publicado: (2025)
ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning
por: Chen, Mingyang, et al.
Publicado: (2025)
por: Chen, Mingyang, et al.
Publicado: (2025)
G1: Teaching LLMs to Reason on Graphs with Reinforcement Learning
por: Guo, Xiaojun, et al.
Publicado: (2025)
por: Guo, Xiaojun, et al.
Publicado: (2025)
SATURN: SAT-based Reinforcement Learning to Unleash LLMs Reasoning
por: Liu, Huanyu, et al.
Publicado: (2025)
por: Liu, Huanyu, et al.
Publicado: (2025)
Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games
por: He, Yidong, et al.
Publicado: (2026)
por: He, Yidong, et al.
Publicado: (2026)
LLMSense: Harnessing LLMs for High-level Reasoning Over Spatiotemporal Sensor Traces
por: Ouyang, Xiaomin, et al.
Publicado: (2024)
por: Ouyang, Xiaomin, et al.
Publicado: (2024)
Toward Scientific Reasoning in LLMs: Training from Expert Discussions via Reinforcement Learning
por: Yin, Ming, et al.
Publicado: (2025)
por: Yin, Ming, et al.
Publicado: (2025)
DRAFT-RL: Multi-Agent Chain-of-Draft Reasoning for Reinforcement Learning-Enhanced LLMs
por: Li, Yuanhao, et al.
Publicado: (2025)
por: Li, Yuanhao, et al.
Publicado: (2025)
Boosting Accuracy and Efficiency of Budget Forcing in LLMs via Reinforcement Learning for Mathematical Reasoning
por: Tarunokusumo, Ravindra Aribowo, et al.
Publicado: (2025)
por: Tarunokusumo, Ravindra Aribowo, et al.
Publicado: (2025)
CosmicFish-HRM: Adaptive Reasoning via Hierarchical Recurrent Mechanisms in Compact Language Models
por: Lakkapragada, Venkat Akhil
Publicado: (2026)
por: Lakkapragada, Venkat Akhil
Publicado: (2026)
Experience as a Compass: Multi-agent RAG with Evolving Orchestration and Agent Prompts
por: Li, Sha, et al.
Publicado: (2026)
por: Li, Sha, et al.
Publicado: (2026)
Human-Inspired Multi-Level Reinforcement Learning
por: Wu, Mingkang, et al.
Publicado: (2025)
por: Wu, Mingkang, et al.
Publicado: (2025)
Human-Inspired Framework to Accelerate Reinforcement Learning
por: Beikmohammadi, Ali, et al.
Publicado: (2023)
por: Beikmohammadi, Ali, et al.
Publicado: (2023)
Emotion-Coherent Reasoning for Multimodal LLMs via Emotional Rationale Verifier
por: Rha, Hyeongseop, et al.
Publicado: (2025)
por: Rha, Hyeongseop, et al.
Publicado: (2025)
Unlocking Reasoning Capabilities in LLMs via Reinforcement Learning Exploration
por: Deng, Wenhao, et al.
Publicado: (2025)
por: Deng, Wenhao, et al.
Publicado: (2025)
Prompted Policy Search: Reinforcement Learning through Linguistic and Numerical Reasoning in LLMs
por: Zhou, Yifan, et al.
Publicado: (2025)
por: Zhou, Yifan, et al.
Publicado: (2025)
Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs
por: Wen, Xumeng, et al.
Publicado: (2025)
por: Wen, Xumeng, et al.
Publicado: (2025)
ARDNS-FN-Quantum: A Quantum-Enhanced Reinforcement Learning Framework with Cognitive-Inspired Adaptive Exploration for Dynamic Environments
por: de Sousa, Umberto Gonçalves
Publicado: (2025)
por: de Sousa, Umberto Gonçalves
Publicado: (2025)
Xiangqi-R1: Enhancing Spatial Strategic Reasoning in LLMs for Chinese Chess via Reinforcement Learning
por: Chen, Yuhao, et al.
Publicado: (2025)
por: Chen, Yuhao, et al.
Publicado: (2025)
Brain-Inspired Planning for Better Generalization in Reinforcement Learning
por: Zhao, Mingde "Harry"
Publicado: (2025)
por: Zhao, Mingde "Harry"
Publicado: (2025)
R3-RAG: Learning Step-by-Step Reasoning and Retrieval for LLMs via Reinforcement Learning
por: Li, Yuan, et al.
Publicado: (2025)
por: Li, Yuan, et al.
Publicado: (2025)
Deep Reinforcement Learning with Gradient Eligibility Traces
por: Elelimy, Esraa, et al.
Publicado: (2025)
por: Elelimy, Esraa, et al.
Publicado: (2025)
Toward IIT-Inspired Consciousness in LLMs: A Reward-Based Learning Framework
por: Akbari, Hamid Reza, et al.
Publicado: (2026)
por: Akbari, Hamid Reza, et al.
Publicado: (2026)
Tracing the Traces: Latent Temporal Signals for Efficient and Accurate Reasoning
por: Vilas, Martina G., et al.
Publicado: (2025)
por: Vilas, Martina G., et al.
Publicado: (2025)
TimeMaster: Training Time-Series Multimodal LLMs to Reason via Reinforcement Learning
por: Zhang, Junru, et al.
Publicado: (2025)
por: Zhang, Junru, et al.
Publicado: (2025)
DiFFPO: Training Diffusion LLMs to Reason Fast and Furious via Reinforcement Learning
por: Zhao, Hanyang, et al.
Publicado: (2025)
por: Zhao, Hanyang, et al.
Publicado: (2025)
Ejemplares similares
-
Hybrid Safety Verification of Multi-Agent Systems using $ψ$-Weighted CBFs and PAC Guarantees
por: Margapuri, Venkat, et al.
Publicado: (2025) -
Diagnosis and Severity Assessment of Ulcerative Colitis using Self Supervised Learning
por: Margapuri, Venkat
Publicado: (2024) -
Predicting Mortality and Functional Status Scores of Traumatic Brain Injury Patients using Supervised Machine Learning
por: Steinmetz, Lucas, et al.
Publicado: (2024) -
Seed Kernel Counting using Domain Randomization and Object Tracking Neural Networks
por: Margapuri, Venkat, et al.
Publicado: (2023) -
Prompt Informed Reinforcement Learning for Visual Coverage Path Planning
por: Margapuri, Venkat
Publicado: (2025)