Human-Inspired Continuous Learning of Internal Reasoning Processes: Learning How to Think for Adaptive AI Systems
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Su, Hong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Human Simulation Computation: A Human-Inspired Framework for Adaptive AI Systems
von: Su, Hong
Veröffentlicht: (2026)
von: Su, Hong
Veröffentlicht: (2026)
Human-Inspired Learning for Large Language Models via Obvious Record and Maximum-Entropy Method Discovery
von: Su, Hong
Veröffentlicht: (2025)
von: Su, Hong
Veröffentlicht: (2025)
Simulating Human Cognition: Heartbeat-Driven Autonomous Thinking Activity Scheduling for LLM-based AI systems
von: Su, Hong
Veröffentlicht: (2026)
von: Su, Hong
Veröffentlicht: (2026)
When to Continue Thinking: Adaptive Thinking Mode Switching for Efficient Reasoning
von: Zhang, Xiaoyun, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoyun, et al.
Veröffentlicht: (2025)
How Clinicians Think and What AI Can Learn From It
von: Sengupta, Dipayan, et al.
Veröffentlicht: (2026)
von: Sengupta, Dipayan, et al.
Veröffentlicht: (2026)
Active Thinking Model: A Goal-Directed Self-Improving Framework for Real-World Adaptive Intelligence
von: Su, Hong
Veröffentlicht: (2025)
von: Su, Hong
Veröffentlicht: (2025)
Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought
von: Xiang, Violet, et al.
Veröffentlicht: (2025)
von: Xiang, Violet, et al.
Veröffentlicht: (2025)
Omni-AutoThink: Adaptive Multimodal Reasoning via Reinforcement Learning
von: Yang, Dongchao, et al.
Veröffentlicht: (2025)
von: Yang, Dongchao, et al.
Veröffentlicht: (2025)
Nice Fold or Hero Call: Learning Budget-Efficient Thinking for Adaptive Reasoning
von: Zhou, Zhaomeng, et al.
Veröffentlicht: (2026)
von: Zhou, Zhaomeng, et al.
Veröffentlicht: (2026)
Disagreements in Reasoning: How a Model's Thinking Process Dictates Persuasion in Multi-Agent Systems
von: Zhao, Haodong, et al.
Veröffentlicht: (2025)
von: Zhao, Haodong, et al.
Veröffentlicht: (2025)
Base Models Know How to Reason, Thinking Models Learn When
von: Venhoff, Constantin, et al.
Veröffentlicht: (2025)
von: Venhoff, Constantin, et al.
Veröffentlicht: (2025)
Reasoning Before Diagnosis: Physician-Inspired Structured Thinking for ECG Classification
von: Wu, Yang, et al.
Veröffentlicht: (2026)
von: Wu, Yang, et al.
Veröffentlicht: (2026)
From <Answer> to <Think>: Multidimensional Supervision of Reasoning Process for LLM Optimization
von: Wang, Beining, et al.
Veröffentlicht: (2025)
von: Wang, Beining, et al.
Veröffentlicht: (2025)
Just Enough Thinking: Efficient Reasoning with Adaptive Length Penalties Reinforcement Learning
von: Xiang, Violet, et al.
Veröffentlicht: (2025)
von: Xiang, Violet, et al.
Veröffentlicht: (2025)
System 2 Reasoning for Human-AI Alignment: Generality and Adaptivity via ARC-AGI
von: Kim, Sejin, et al.
Veröffentlicht: (2024)
von: Kim, Sejin, et al.
Veröffentlicht: (2024)
Learning How Hard to Think: Input-Adaptive Allocation of LM Computation
von: Damani, Mehul, et al.
Veröffentlicht: (2024)
von: Damani, Mehul, et al.
Veröffentlicht: (2024)
Teaching AI to Remember: Insights from Brain-Inspired Replay in Continual Learning
von: Kim, Jina
Veröffentlicht: (2025)
von: Kim, Jina
Veröffentlicht: (2025)
NeuroHex: A Brain-Inspired Hex Coordinate System to Enable Highly Computationally-Efficient World Models for Continuous Online-Adaptive Learning
von: Jacobson, Quinn, et al.
Veröffentlicht: (2026)
von: Jacobson, Quinn, et al.
Veröffentlicht: (2026)
Think How to Think: Mitigating Overthinking with Autonomous Difficulty Cognition in Large Reasoning Models
von: Liu, Yongjiang, et al.
Veröffentlicht: (2025)
von: Liu, Yongjiang, et al.
Veröffentlicht: (2025)
Adaptive Dual Reasoner: Large Reasoning Models Can Think Efficiently by Hybrid Reasoning
von: Zhang, Yujian, et al.
Veröffentlicht: (2025)
von: Zhang, Yujian, et al.
Veröffentlicht: (2025)
Autonomous Question Formation for Large Language Model-Driven AI Systems
von: Su, Hong
Veröffentlicht: (2026)
von: Su, Hong
Veröffentlicht: (2026)
HiCL: Hippocampal-Inspired Continual Learning
von: Kapoor, Kushal, et al.
Veröffentlicht: (2025)
von: Kapoor, Kushal, et al.
Veröffentlicht: (2025)
Learning to Solve Geometry Problems via Simulating Human Dual-Reasoning Process
von: Xiao, Tong, et al.
Veröffentlicht: (2024)
von: Xiao, Tong, et al.
Veröffentlicht: (2024)
Thought Cloning: Learning to Think while Acting by Imitating Human Thinking
von: Hu, Shengran, et al.
Veröffentlicht: (2023)
von: Hu, Shengran, et al.
Veröffentlicht: (2023)
Think Smarter not Harder: Adaptive Reasoning with Inference Aware Optimization
von: Yu, Zishun, et al.
Veröffentlicht: (2025)
von: Yu, Zishun, et al.
Veröffentlicht: (2025)
Internalizing Outcome Supervision into Process Supervision: A New Paradigm for Reinforcement Learning for Reasoning
von: Ding, Fei, et al.
Veröffentlicht: (2026)
von: Ding, Fei, et al.
Veröffentlicht: (2026)
Human-Inspired Framework to Accelerate Reinforcement Learning
von: Beikmohammadi, Ali, et al.
Veröffentlicht: (2023)
von: Beikmohammadi, Ali, et al.
Veröffentlicht: (2023)
Human-Inspired Multi-Level Reinforcement Learning
von: Wu, Mingkang, et al.
Veröffentlicht: (2025)
von: Wu, Mingkang, et al.
Veröffentlicht: (2025)
Think in Games: Learning to Reason in Games via Reinforcement Learning with Large Language Models
von: Liao, Yi, et al.
Veröffentlicht: (2025)
von: Liao, Yi, et al.
Veröffentlicht: (2025)
How To Think About End-To-End Encryption and AI: Training, Processing, Disclosure, and Consent
von: Knodel, Mallory, et al.
Veröffentlicht: (2024)
von: Knodel, Mallory, et al.
Veröffentlicht: (2024)
InqEduAgent: Adaptive AI Learning Partners with Gaussian Process Augmentation
von: Yang, Wen-Xi, et al.
Veröffentlicht: (2025)
von: Yang, Wen-Xi, et al.
Veröffentlicht: (2025)
AdaptThink: Reasoning Models Can Learn When to Think
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2025)
Adaptive Collaboration with Humans: Metacognitive Policy Optimization for Multi-Agent LLMs with Continual Learning
von: Yang, Wei, et al.
Veröffentlicht: (2026)
von: Yang, Wei, et al.
Veröffentlicht: (2026)
Method-Based Reasoning for Large Language Models: Extraction, Reuse, and Continuous Improvement
von: Su, Hong
Veröffentlicht: (2025)
von: Su, Hong
Veröffentlicht: (2025)
PEPS: Quantum-Inspired Reinforcement Learning for Coherent Reasoning Traces in LLMs
von: Margapuri, Venkat, et al.
Veröffentlicht: (2025)
von: Margapuri, Venkat, et al.
Veröffentlicht: (2025)
Personalized Artificial General Intelligence (AGI) via Neuroscience-Inspired Continuous Learning Systems
von: Gupta, Rajeev, et al.
Veröffentlicht: (2025)
von: Gupta, Rajeev, et al.
Veröffentlicht: (2025)
AdapThink: Adaptive Thinking Preferences for Reasoning Language Model
von: Wan, Xu, et al.
Veröffentlicht: (2025)
von: Wan, Xu, et al.
Veröffentlicht: (2025)
CoT-Space: A Theoretical Framework for Internal Slow-Thinking via Reinforcement Learning
von: Gan, Zeyu, et al.
Veröffentlicht: (2025)
von: Gan, Zeyu, et al.
Veröffentlicht: (2025)
Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces
von: Su, DiJia, et al.
Veröffentlicht: (2024)
von: Su, DiJia, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Human Simulation Computation: A Human-Inspired Framework for Adaptive AI Systems
von: Su, Hong
Veröffentlicht: (2026) -
Human-Inspired Learning for Large Language Models via Obvious Record and Maximum-Entropy Method Discovery
von: Su, Hong
Veröffentlicht: (2025) -
Simulating Human Cognition: Heartbeat-Driven Autonomous Thinking Activity Scheduling for LLM-based AI systems
von: Su, Hong
Veröffentlicht: (2026) -
When to Continue Thinking: Adaptive Thinking Mode Switching for Efficient Reasoning
von: Zhang, Xiaoyun, et al.
Veröffentlicht: (2025) -
How Clinicians Think and What AI Can Learn From It
von: Sengupta, Dipayan, et al.
Veröffentlicht: (2026)