Thinking Fast and Slow with Deep Learning and Tree Search
Fuente:
arXiv
Saved in:
| Main Authors: | Anthony, Thomas, Tian, Zheng, Barber, David |
|---|---|
| Format: | Preprint |
| Published: |
2017
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces
by: Su, DiJia, et al.
Published: (2024)
by: Su, DiJia, et al.
Published: (2024)
DSADF: Thinking Fast and Slow for Decision Making
by: Dou, Zhihao, et al.
Published: (2025)
by: Dou, Zhihao, et al.
Published: (2025)
Thinking While Listening: Fast-Slow Recurrence for Long-Horizon Sequential Modeling
by: Takashiro, Shota, et al.
Published: (2026)
by: Takashiro, Shota, et al.
Published: (2026)
Agents Thinking Fast and Slow: A Talker-Reasoner Architecture
by: Christakopoulou, Konstantina, et al.
Published: (2024)
by: Christakopoulou, Konstantina, et al.
Published: (2024)
Thinker: Learning to Think Fast and Slow
by: Chung, Stephen, et al.
Published: (2025)
by: Chung, Stephen, et al.
Published: (2025)
DAST: Difficulty-Adaptive Slow-Thinking for Large Reasoning Models
by: Shen, Yi, et al.
Published: (2025)
by: Shen, Yi, et al.
Published: (2025)
What Happened in LLMs Layers when Trained for Fast vs. Slow Thinking: A Gradient Perspective
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
Time Series Forecasting as Reasoning: A Slow-Thinking Approach with Reinforced LLMs
by: Zhou, Yitong, et al.
Published: (2025)
by: Zhou, Yitong, et al.
Published: (2025)
ThinkGuard: Deliberative Slow Thinking Leads to Cautious Guardrails
by: Wen, Xiaofei, et al.
Published: (2025)
by: Wen, Xiaofei, et al.
Published: (2025)
Slow-Fast Policy Optimization: Reposition-Before-Update for LLM Reasoning
by: Wang, Ziyan, et al.
Published: (2025)
by: Wang, Ziyan, et al.
Published: (2025)
Fast-Slow Co-advancing Optimizer: Toward Harmonious Adversarial Training of GAN
by: Wang, Lin, et al.
Published: (2025)
by: Wang, Lin, et al.
Published: (2025)
Reactive Diffusion Policy: Slow-Fast Visual-Tactile Policy Learning for Contact-Rich Manipulation
by: Xue, Han, et al.
Published: (2025)
by: Xue, Han, et al.
Published: (2025)
ImplicitRDP: An End-to-End Visual-Force Diffusion Policy with Structural Slow-Fast Learning
by: Chen, Wendi, et al.
Published: (2025)
by: Chen, Wendi, et al.
Published: (2025)
System-1.x: Learning to Balance Fast and Slow Planning with Language Models
by: Saha, Swarnadeep, et al.
Published: (2024)
by: Saha, Swarnadeep, et al.
Published: (2024)
Deep Causal Learning: Representation, Discovery and Inference
by: Deng, Zizhen, et al.
Published: (2022)
by: Deng, Zizhen, et al.
Published: (2022)
Fast Value Tracking for Deep Reinforcement Learning
by: Shih, Frank, et al.
Published: (2024)
by: Shih, Frank, et al.
Published: (2024)
Multi-Objective Neural Architecture Search by Learning Search Space Partitions
by: Zhao, Yiyang, et al.
Published: (2024)
by: Zhao, Yiyang, et al.
Published: (2024)
SlowFast-VGen: Slow-Fast Learning for Action-Driven Long Video Generation
by: Hong, Yining, et al.
Published: (2024)
by: Hong, Yining, et al.
Published: (2024)
STREAM-VAE: Dual-Path Routing for Slow and Fast Dynamics in Vehicle Telemetry Anomaly Detection
by: Özer, Kadir-Kaan, et al.
Published: (2025)
by: Özer, Kadir-Kaan, et al.
Published: (2025)
Slow-Fast Inference: Training-Free Inference Acceleration via Within-Sentence Support Stability
by: Xie, Xingyu, et al.
Published: (2026)
by: Xie, Xingyu, et al.
Published: (2026)
ePC: Fast and Deep Predictive Coding for Digital Hardware
by: Goemaere, Cédric, et al.
Published: (2025)
by: Goemaere, Cédric, et al.
Published: (2025)
Explaining, Fast and Slow: Abstraction and Refinement of Provable Explanations
by: Bassan, Shahaf, et al.
Published: (2025)
by: Bassan, Shahaf, et al.
Published: (2025)
Learn to Think: Bootstrapping LLM Reasoning Capability Through Graph Representation Learning
by: Gao, Hang, et al.
Published: (2025)
by: Gao, Hang, et al.
Published: (2025)
Online DPO: Online Direct Preference Optimization with Fast-Slow Chasing
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Emergent Slow Thinking in LLMs as Inverse Tree Freezing
by: Hu, Sihan, et al.
Published: (2025)
by: Hu, Sihan, et al.
Published: (2025)
Thought Cloning: Learning to Think while Acting by Imitating Human Thinking
by: Hu, Shengran, et al.
Published: (2023)
by: Hu, Shengran, et al.
Published: (2023)
Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning
by: Xie, Yuxi, et al.
Published: (2024)
by: Xie, Yuxi, et al.
Published: (2024)
Tree Search for LLM Agent Reinforcement Learning
by: Ji, Yuxiang, et al.
Published: (2025)
by: Ji, Yuxiang, et al.
Published: (2025)
Tail-Risk-Safe Monte Carlo Tree Search under PAC-Level Guarantees
by: Zhang, Zuyuan, et al.
Published: (2025)
by: Zhang, Zuyuan, et al.
Published: (2025)
A* Search Without Expansions: Learning Heuristic Functions with Deep Q-Networks
by: Agostinelli, Forest, et al.
Published: (2021)
by: Agostinelli, Forest, et al.
Published: (2021)
Online Control of Adaptive Large Neighborhood Search using Deep Reinforcement Learning
by: Reijnen, Robbert, et al.
Published: (2022)
by: Reijnen, Robbert, et al.
Published: (2022)
Think in Blocks: Adaptive Reasoning from Direct Response to Deep Reasoning
by: Zhu, Yekun, et al.
Published: (2025)
by: Zhu, Yekun, et al.
Published: (2025)
Accelerating Diffusion Large Language Models with SlowFast Sampling: The Three Golden Principles
by: Wei, Qingyan, et al.
Published: (2025)
by: Wei, Qingyan, et al.
Published: (2025)
On the Emergence of Thinking in LLMs I: Searching for the Right Intuition
by: Ye, Guanghao, et al.
Published: (2025)
by: Ye, Guanghao, et al.
Published: (2025)
The DeepXube Software Package for Solving Pathfinding Problems with Learned Heuristic Functions and Search
by: Agostinelli, Forest
Published: (2026)
by: Agostinelli, Forest
Published: (2026)
Fast and Robust Likelihood-Guided Diffusion Posterior Sampling with Amortized Variational Inference
by: Zheng, Léon, et al.
Published: (2026)
by: Zheng, Léon, et al.
Published: (2026)
LiteSearch: Efficacious Tree Search for LLM
by: Wang, Ante, et al.
Published: (2024)
by: Wang, Ante, et al.
Published: (2024)
Multimodal Deep Learning for Low-Resource Settings: A Vector Embedding Alignment Approach for Healthcare Applications
by: Restrepo, David, et al.
Published: (2024)
by: Restrepo, David, et al.
Published: (2024)
Fast and Exact Enumeration of Deep Networks Partitions Regions
by: Balestriero, Randall, et al.
Published: (2024)
by: Balestriero, Randall, et al.
Published: (2024)
Fast Unsupervised Deep Outlier Model Selection with Hypernetworks
by: Ding, Xueying, et al.
Published: (2023)
by: Ding, Xueying, et al.
Published: (2023)
Similar Items
-
Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces
by: Su, DiJia, et al.
Published: (2024) -
DSADF: Thinking Fast and Slow for Decision Making
by: Dou, Zhihao, et al.
Published: (2025) -
Thinking While Listening: Fast-Slow Recurrence for Long-Horizon Sequential Modeling
by: Takashiro, Shota, et al.
Published: (2026) -
Agents Thinking Fast and Slow: A Talker-Reasoner Architecture
by: Christakopoulou, Konstantina, et al.
Published: (2024) -
Thinker: Learning to Think Fast and Slow
by: Chung, Stephen, et al.
Published: (2025)