Thought Cloning: Learning to Think while Acting by Imitating Human Thinking
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Shengran, Clune, Jeff |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Automated Capability Discovery via Foundation Model Self-Exploration
by: Lu, Cong, et al.
Published: (2025)
by: Lu, Cong, et al.
Published: (2025)
Intelligent Go-Explore: Standing on the Shoulders of Giant Foundation Models
by: Lu, Cong, et al.
Published: (2024)
by: Lu, Cong, et al.
Published: (2024)
Learning to Continually Learn via Meta-learning Agentic Memory Designs
by: Xiong, Yiming, et al.
Published: (2026)
by: Xiong, Yiming, et al.
Published: (2026)
Automated Design of Agentic Systems
by: Hu, Shengran, et al.
Published: (2024)
by: Hu, Shengran, et al.
Published: (2024)
First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs
by: Norman, Ben, et al.
Published: (2023)
by: Norman, Ben, et al.
Published: (2023)
The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
by: Yamada, Yutaro, et al.
Published: (2025)
by: Yamada, Yutaro, et al.
Published: (2025)
Foundation Model Self-Play: Open-Ended Strategy Innovation via Foundation Models
by: Dharna, Aaron, et al.
Published: (2025)
by: Dharna, Aaron, et al.
Published: (2025)
Dyna-Think: Synergizing Reasoning, Acting, and World Model Simulation in AI Agents
by: Yu, Xiao, et al.
Published: (2025)
by: Yu, Xiao, et al.
Published: (2025)
Think When You Need: Self-Adaptive Chain-of-Thought Learning
by: Yang, Junjie, et al.
Published: (2025)
by: Yang, Junjie, et al.
Published: (2025)
OMNI: Open-endedness via Models of human Notions of Interestingness
by: Zhang, Jenny, et al.
Published: (2023)
by: Zhang, Jenny, et al.
Published: (2023)
Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents
by: Zhang, Jenny, et al.
Published: (2025)
by: Zhang, Jenny, et al.
Published: (2025)
Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning
by: Foster, Dylan J., et al.
Published: (2024)
by: Foster, Dylan J., et al.
Published: (2024)
AdaptThink: Reasoning Models Can Learn When to Think
by: Zhang, Jiajie, et al.
Published: (2025)
by: Zhang, Jiajie, et al.
Published: (2025)
Learning Soft Driving Constraints from Vectorized Scene Embeddings while Imitating Expert Trajectories
by: Mobarakeh, Niloufar Saeidi, et al.
Published: (2024)
by: Mobarakeh, Niloufar Saeidi, et al.
Published: (2024)
GBC: Generalized Behavior-Cloning Framework for Whole-Body Humanoid Imitation
by: Yao, Yifei, et al.
Published: (2025)
by: Yao, Yifei, et al.
Published: (2025)
Thinking Preference Optimization
by: Yang, Wang, et al.
Published: (2025)
by: Yang, Wang, et al.
Published: (2025)
Think Consistently, Reason Efficiently: Energy-Based Calibration for Implicit Chain-of-Thought
by: Chen, Zhikang, et al.
Published: (2025)
by: Chen, Zhikang, et al.
Published: (2025)
Demystifying Hybrid Thinking: Can LLMs Truly Switch Between Think and No-Think?
by: Wang, Shouren, et al.
Published: (2025)
by: Wang, Shouren, et al.
Published: (2025)
What LLMs Think When You Don't Tell Them What to Think About?
by: Kwon, Yongchan, et al.
Published: (2026)
by: Kwon, Yongchan, et al.
Published: (2026)
Thinking Fast and Slow with Deep Learning and Tree Search
by: Anthony, Thomas, et al.
Published: (2017)
by: Anthony, Thomas, et al.
Published: (2017)
Think Right: Learning to Mitigate Under-Over Thinking via Adaptive, Attentive Compression
by: Singh, Joykirat, et al.
Published: (2025)
by: Singh, Joykirat, et al.
Published: (2025)
Continual learning under domain transfer with sparse synaptic bursting
by: Beaulieu, Shawn L., et al.
Published: (2021)
by: Beaulieu, Shawn L., et al.
Published: (2021)
Imitation Bootstrapped Reinforcement Learning
by: Hu, Hengyuan, et al.
Published: (2023)
by: Hu, Hengyuan, et al.
Published: (2023)
SALT: Steering Activations towards Leakage-free Thinking in Chain of Thought
by: Batra, Shourya, et al.
Published: (2025)
by: Batra, Shourya, et al.
Published: (2025)
Learning to Think from Multiple Thinkers
by: Joshi, Nirmit, et al.
Published: (2026)
by: Joshi, Nirmit, et al.
Published: (2026)
Imitation Learning for Multi-turn LM Agents via On-policy Expert Corrections
by: Lauffer, Niklas, et al.
Published: (2025)
by: Lauffer, Niklas, et al.
Published: (2025)
To Think or Not to Think: Exploring the Unthinking Vulnerability in Large Reasoning Models
by: Zhu, Zihao, et al.
Published: (2025)
by: Zhu, Zihao, et al.
Published: (2025)
AdapThink: Adaptive Thinking Preferences for Reasoning Language Model
by: Wan, Xu, et al.
Published: (2025)
by: Wan, Xu, et al.
Published: (2025)
Learn to Think: Bootstrapping LLM Reasoning Capability Through Graph Representation Learning
by: Gao, Hang, et al.
Published: (2025)
by: Gao, Hang, et al.
Published: (2025)
Nemotron-CrossThink: Scaling Self-Learning beyond Math Reasoning
by: Akter, Syeda Nahida, et al.
Published: (2025)
by: Akter, Syeda Nahida, et al.
Published: (2025)
Base Models Know How to Reason, Thinking Models Learn When
by: Venhoff, Constantin, et al.
Published: (2025)
by: Venhoff, Constantin, et al.
Published: (2025)
To Think or Not to Think: The Hidden Cost of Meta-Training with Excessive CoT Examples
by: Kothapalli, Vignesh, et al.
Published: (2025)
by: Kothapalli, Vignesh, et al.
Published: (2025)
Let Me Think! A Long Chain-of-Thought Can Be Worth Exponentially Many Short Ones
by: Mirtaheri, Parsa, et al.
Published: (2025)
by: Mirtaheri, Parsa, et al.
Published: (2025)
Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse
by: Liu, Ryan, et al.
Published: (2024)
by: Liu, Ryan, et al.
Published: (2024)
Brain-Inspired Two-Stage Approach: Enhancing Mathematical Reasoning by Imitating Human Thought Processes
by: Chen, Yezeng, et al.
Published: (2024)
by: Chen, Yezeng, et al.
Published: (2024)
ResNets Are Deeper Than You Think
by: Mehmeti-Göpel, Christian H. X. Ali, et al.
Published: (2025)
by: Mehmeti-Göpel, Christian H. X. Ali, et al.
Published: (2025)
Rethinking Thinking Tokens: LLMs as Improvement Operators
by: Madaan, Lovish, et al.
Published: (2025)
by: Madaan, Lovish, et al.
Published: (2025)
DSADF: Thinking Fast and Slow for Decision Making
by: Dou, Zhihao, et al.
Published: (2025)
by: Dou, Zhihao, et al.
Published: (2025)
Building Machines that Learn and Think with People
by: Collins, Katherine M., et al.
Published: (2024)
by: Collins, Katherine M., et al.
Published: (2024)
Just Enough Thinking: Efficient Reasoning with Adaptive Length Penalties Reinforcement Learning
by: Xiang, Violet, et al.
Published: (2025)
by: Xiang, Violet, et al.
Published: (2025)
Similar Items
-
Automated Capability Discovery via Foundation Model Self-Exploration
by: Lu, Cong, et al.
Published: (2025) -
Intelligent Go-Explore: Standing on the Shoulders of Giant Foundation Models
by: Lu, Cong, et al.
Published: (2024) -
Learning to Continually Learn via Meta-learning Agentic Memory Designs
by: Xiong, Yiming, et al.
Published: (2026) -
Automated Design of Agentic Systems
by: Hu, Shengran, et al.
Published: (2024) -
First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs
by: Norman, Ben, et al.
Published: (2023)