Beyond Fast and Slow: Cognitive-Inspired Elastic Reasoning for Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Jinwu, Yang, Dongjin, Bian, Langyu, Wen, Zhiquan, Wang, Yufeng, Chen, Yaofo, Xiao, Bin, Li, Yuanqing, Tan, Mingkui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Model Scaling: Test-Time Intervention for Efficient Deep Reasoning
by: Wang, Qianyue, et al.
Published: (2026)
by: Wang, Qianyue, et al.
Published: (2026)
Precedent-Informed Reasoning: Mitigating Overthinking in Large Reasoning Models via Test-Time Precedent Learning
by: Wang, Qianyue, et al.
Published: (2026)
by: Wang, Qianyue, et al.
Published: (2026)
Dynamic Compressing Prompts for Efficient Inference of Large Language Models
by: Hu, Jinwu, et al.
Published: (2025)
by: Hu, Jinwu, et al.
Published: (2025)
Test-Time Learning for Large Language Models
by: Hu, Jinwu, et al.
Published: (2025)
by: Hu, Jinwu, et al.
Published: (2025)
Curse of High Dimensionality Issue in Transformer for Long-context Modeling
by: Zhang, Shuhai, et al.
Published: (2025)
by: Zhang, Shuhai, et al.
Published: (2025)
Continual Knowledge Adaptation for Reinforcement Learning
by: Hu, Jinwu, et al.
Published: (2025)
by: Hu, Jinwu, et al.
Published: (2025)
Towards Robust and Efficient Cloud-Edge Elastic Model Adaptation via Selective Entropy Distillation
by: Chen, Yaofo, et al.
Published: (2024)
by: Chen, Yaofo, et al.
Published: (2024)
Towards Long Video Understanding via Fine-detailed Video Story Generation
by: You, Zeng, et al.
Published: (2024)
by: You, Zeng, et al.
Published: (2024)
Efficient Dynamic Ensembling for Multiple LLM Experts
by: Hu, Jinwu, et al.
Published: (2024)
by: Hu, Jinwu, et al.
Published: (2024)
Enhancing Perception Capabilities of Multimodal LLMs with Training-Free Fusion
by: Chen, Zhuokun, et al.
Published: (2024)
by: Chen, Zhuokun, et al.
Published: (2024)
Generating Long-form Story Using Dynamic Hierarchical Outlining with Memory-Enhancement
by: Wang, Qianyue, et al.
Published: (2024)
by: Wang, Qianyue, et al.
Published: (2024)
Fast-Slow Thinking GRPO for Large Vision-Language Model Reasoning
by: Xiao, Wenyi, et al.
Published: (2025)
by: Xiao, Wenyi, et al.
Published: (2025)
Core Context Aware Transformers for Long Context Language Modeling
by: Chen, Yaofo, et al.
Published: (2024)
by: Chen, Yaofo, et al.
Published: (2024)
Sensitivity-Aware Post-Training Quantization for Deep Neural Networks
by: Zheng, Zekang, et al.
Published: (2025)
by: Zheng, Zekang, et al.
Published: (2025)
Self-Supervised On-Policy Distillation for Reasoning Language Models
by: Tan, Zhiquan, et al.
Published: (2026)
by: Tan, Zhiquan, et al.
Published: (2026)
Adapt in the Wild: Test-Time Entropy Minimization with Sharpness and Feature Regularization
by: Niu, Shuaicheng, et al.
Published: (2025)
by: Niu, Shuaicheng, et al.
Published: (2025)
ProtoDCS: Towards Robust and Efficient Open-Set Test-Time Adaptation for Vision-Language Models
by: Luo, Wei, et al.
Published: (2026)
by: Luo, Wei, et al.
Published: (2026)
Uncertainty-Calibrated Test-Time Model Adaptation without Forgetting
by: Tan, Mingkui, et al.
Published: (2024)
by: Tan, Mingkui, et al.
Published: (2024)
Inference-Cost-Aware Dynamic Tree Construction for Efficient Inference in Large Language Models
by: Hong, Yinrong, et al.
Published: (2025)
by: Hong, Yinrong, et al.
Published: (2025)
EvidFuse: Writing-Time Evidence Learning for Consistent Text-Chart Data Reporting
by: Lin, Huanxiang, et al.
Published: (2026)
by: Lin, Huanxiang, et al.
Published: (2026)
Zero-Shot Skeleton-Based Action Recognition With Prototype-Guided Feature Alignment
by: Zhou, Kai, et al.
Published: (2025)
by: Zhou, Kai, et al.
Published: (2025)
The Information of Large Language Model Geometry
by: Tan, Zhiquan, et al.
Published: (2024)
by: Tan, Zhiquan, et al.
Published: (2024)
Principal Trotter Observation Error with Truncated Commutators
by: Li, Langyu
Published: (2024)
by: Li, Langyu
Published: (2024)
Instance-level Visual Active Tracking with Occlusion-Aware Planning
by: Sun, Haowei, et al.
Published: (2026)
by: Sun, Haowei, et al.
Published: (2026)
Latent-Condensed Transformer for Efficient Long Context Modeling
by: You, Zeng, et al.
Published: (2026)
by: You, Zeng, et al.
Published: (2026)
Streaming, Fast and Slow: Cognitive Load-Aware Streaming for Efficient LLM Serving
by: Xiao, Chang, et al.
Published: (2025)
by: Xiao, Chang, et al.
Published: (2025)
Training-free Context-adaptive Attention for Efficient Long Context Modeling
by: You, Zeng, et al.
Published: (2025)
by: You, Zeng, et al.
Published: (2025)
Cognitive Decision Routing in Large Language Models: When to Think Fast, When to Think Slow
by: Du, Y., et al.
Published: (2025)
by: Du, Y., et al.
Published: (2025)
Understanding Emotional Body Expressions via Large Language Models
by: Lu, Haifeng, et al.
Published: (2024)
by: Lu, Haifeng, et al.
Published: (2024)
Fast Numerical Solver of Ising Optimization Problems via Pruning and Domain Selection
by: Li, Langyu, et al.
Published: (2023)
by: Li, Langyu, et al.
Published: (2023)
Open-World Drone Active Tracking with Goal-Centered Rewards
by: Sun, Haowei, et al.
Published: (2024)
by: Sun, Haowei, et al.
Published: (2024)
Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning
by: Chen, Hao, et al.
Published: (2025)
by: Chen, Hao, et al.
Published: (2025)
MixRea: Benchmarking Explicit-Implicit Reasoning in Large Language Models
by: Cai, Yuanqing, et al.
Published: (2026)
by: Cai, Yuanqing, et al.
Published: (2026)
Detecting Machine-Generated Texts by Multi-Population Aware Optimization for Maximum Mean Discrepancy
by: Zhang, Shuhai, et al.
Published: (2024)
by: Zhang, Shuhai, et al.
Published: (2024)
A Theoretical Lens for RL-Tuned Language Models via Energy-Based Models
by: Tan, Zhiquan, et al.
Published: (2025)
by: Tan, Zhiquan, et al.
Published: (2025)
MoTVLA: A Vision-Language-Action Model with Unified Fast-Slow Reasoning
by: Huang, Wenhui, et al.
Published: (2025)
by: Huang, Wenhui, et al.
Published: (2025)
Enhancing User-Oriented Proactivity in Open-Domain Dialogues with Critic Guidance
by: Wang, Yufeng, et al.
Published: (2025)
by: Wang, Yufeng, et al.
Published: (2025)
Automated Dominative Subspace Mining for Efficient Neural Architecture Search
by: Chen, Yaofo, et al.
Published: (2022)
by: Chen, Yaofo, et al.
Published: (2022)
PAINT: Partial-Solution Adaptive Interpolated Training for Self-Distilled Reasoners
by: Tan, Zhiquan, et al.
Published: (2026)
by: Tan, Zhiquan, et al.
Published: (2026)
Fast-Slow Efficient Training for Multimodal Large Language Models via Visual Token Pruning
by: Zhang, Dingkun, et al.
Published: (2026)
by: Zhang, Dingkun, et al.
Published: (2026)
Similar Items
-
Beyond Model Scaling: Test-Time Intervention for Efficient Deep Reasoning
by: Wang, Qianyue, et al.
Published: (2026) -
Precedent-Informed Reasoning: Mitigating Overthinking in Large Reasoning Models via Test-Time Precedent Learning
by: Wang, Qianyue, et al.
Published: (2026) -
Dynamic Compressing Prompts for Efficient Inference of Large Language Models
by: Hu, Jinwu, et al.
Published: (2025) -
Test-Time Learning for Large Language Models
by: Hu, Jinwu, et al.
Published: (2025) -
Curse of High Dimensionality Issue in Transformer for Long-context Modeling
by: Zhang, Shuhai, et al.
Published: (2025)