Gespeichert in:
| Hauptverfasser: | Zhu, Fangwei, Sui, Zhifang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.04246 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Chain-of-Thought Tokens are Computer Program Variables
von: Zhu, Fangwei, et al.
Veröffentlicht: (2025)
von: Zhu, Fangwei, et al.
Veröffentlicht: (2025)
Reducing Hallucinations in Entity Abstract Summarization with Facts-Template Decomposition
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
Language Models Encode the Value of Numbers Linearly
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
RICo: Refined In-Context Contribution for Automatic Instruction-Tuning Data Selection
von: Yang, Yixin, et al.
Veröffentlicht: (2025)
von: Yang, Yixin, et al.
Veröffentlicht: (2025)
Chip-Tuning: Classify Before Language Models Say
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)
SCRIBE: Structured Chain Reasoning for Interactive Behaviour Explanations using Tool Calling
von: Fawzi, Fares, et al.
Veröffentlicht: (2025)
von: Fawzi, Fares, et al.
Veröffentlicht: (2025)
CoLT: The conditional localization test for assessing the accuracy of neural posterior estimates
von: Chen, Tianyu, et al.
Veröffentlicht: (2025)
von: Chen, Tianyu, et al.
Veröffentlicht: (2025)
Chain-of-Tools: Utilizing Massive Unseen Tools in the CoT Reasoning of Frozen Language Models
von: Wu, Mengsong, et al.
Veröffentlicht: (2025)
von: Wu, Mengsong, et al.
Veröffentlicht: (2025)
Latent Preference Modeling for Cross-Session Personalized Tool Calling
von: Yoon, Yejin, et al.
Veröffentlicht: (2026)
von: Yoon, Yejin, et al.
Veröffentlicht: (2026)
CoMet: Metaphor-Driven Covert Communication for Multi-Agent Language Games
von: Xu, Shuhang, et al.
Veröffentlicht: (2025)
von: Xu, Shuhang, et al.
Veröffentlicht: (2025)
SCoRE: Benchmarking Long-Chain Reasoning in Commonsense Scenarios
von: Zhan, Weidong, et al.
Veröffentlicht: (2025)
von: Zhan, Weidong, et al.
Veröffentlicht: (2025)
LLM Latent Reasoning as Chain of Superposition
von: Deng, Jingcheng, et al.
Veröffentlicht: (2025)
von: Deng, Jingcheng, et al.
Veröffentlicht: (2025)
Exploring Activation Patterns of Parameters in Language Models
von: Wang, Yudong, et al.
Veröffentlicht: (2024)
von: Wang, Yudong, et al.
Veröffentlicht: (2024)
CoUDA: Coherence Evaluation via Unified Data Augmentation
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
von: Zhu, Dawei, et al.
Veröffentlicht: (2024)
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
von: Li, Zheng, et al.
Veröffentlicht: (2025)
von: Li, Zheng, et al.
Veröffentlicht: (2025)
How Far are LLMs from Being Our Digital Twins? A Benchmark for Persona-Based Behavior Chain Simulation
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
When2Call: When (not) to Call Tools
von: Ross, Hayley, et al.
Veröffentlicht: (2025)
von: Ross, Hayley, et al.
Veröffentlicht: (2025)
Reinforced Context Order Recovery for Adaptive Reasoning and Planning
von: Ma, Long, et al.
Veröffentlicht: (2025)
von: Ma, Long, et al.
Veröffentlicht: (2025)
HistLens: Mapping Idea Change across Concepts and Corpora
von: Jing, Yi, et al.
Veröffentlicht: (2026)
von: Jing, Yi, et al.
Veröffentlicht: (2026)
Towards Better RL Training Data Utilization via Second-Order Rollout
von: Yang, Zhe, et al.
Veröffentlicht: (2026)
von: Yang, Zhe, et al.
Veröffentlicht: (2026)
Not All Demonstration Examples are Equally Beneficial: Reweighting Demonstration Examples for In-Context Learning
von: Yang, Zhe, et al.
Veröffentlicht: (2023)
von: Yang, Zhe, et al.
Veröffentlicht: (2023)
Latent Chain-of-Thought for Visual Reasoning
von: Sun, Guohao, et al.
Veröffentlicht: (2025)
von: Sun, Guohao, et al.
Veröffentlicht: (2025)
Efficient Tool Use with Chain-of-Abstraction Reasoning
von: Gao, Silin, et al.
Veröffentlicht: (2024)
von: Gao, Silin, et al.
Veröffentlicht: (2024)
L2V-CoT: Cross-Modal Transfer of Chain-of-Thought Reasoning via Latent Intervention
von: Zhan, Yuliang, et al.
Veröffentlicht: (2025)
von: Zhan, Yuliang, et al.
Veröffentlicht: (2025)
LLM Agents Already Know When to Call Tools -- Even Without Reasoning
von: Sun, Chung-En, et al.
Veröffentlicht: (2026)
von: Sun, Chung-En, et al.
Veröffentlicht: (2026)
CoT-Evo: Evolutionary Distillation of Chain-of-Thought for Scientific Reasoning
von: Feng, Kehua, et al.
Veröffentlicht: (2025)
von: Feng, Kehua, et al.
Veröffentlicht: (2025)
Alignment for Efficient Tool Calling of Large Language Models
von: Xu, Hongshen, et al.
Veröffentlicht: (2025)
von: Xu, Hongshen, et al.
Veröffentlicht: (2025)
LaRS: Latent Reasoning Skills for Chain-of-Thought Reasoning
von: Xu, Zifan, et al.
Veröffentlicht: (2023)
von: Xu, Zifan, et al.
Veröffentlicht: (2023)
From Mathematical Reasoning to Code: Generalization of Process Reward Models in Test-Time Scaling
von: Chen, Zhengyu, et al.
Veröffentlicht: (2025)
von: Chen, Zhengyu, et al.
Veröffentlicht: (2025)
Reasoning Beyond Language: A Comprehensive Survey on Latent Chain-of-Thought Reasoning
von: Chen, Xinghao, et al.
Veröffentlicht: (2025)
von: Chen, Xinghao, et al.
Veröffentlicht: (2025)
Selective Latent Thinking: Adaptive Compression of LLM Reasoning Chains
von: Xie, Hui, et al.
Veröffentlicht: (2026)
von: Xie, Hui, et al.
Veröffentlicht: (2026)
PruneTIR: Inference-Time Tool Call Pruning for Effective yet Efficient Tool-Integrated Reasoning
von: Zhang, Luan, et al.
Veröffentlicht: (2026)
von: Zhang, Luan, et al.
Veröffentlicht: (2026)
HauntAttack: When Attack Follows Reasoning as a Shadow
von: Ma, Jingyuan, et al.
Veröffentlicht: (2025)
von: Ma, Jingyuan, et al.
Veröffentlicht: (2025)
Plug-and-Play Training Framework for Preference Optimization
von: Ma, Jingyuan, et al.
Veröffentlicht: (2024)
von: Ma, Jingyuan, et al.
Veröffentlicht: (2024)
Enhancing Reliability across Short and Long-Form QA via Reinforcement Learning
von: Wang, Yudong, et al.
Veröffentlicht: (2025)
von: Wang, Yudong, et al.
Veröffentlicht: (2025)
Can Large Multimodal Models Uncover Deep Semantics Behind Images?
von: Yang, Yixin, et al.
Veröffentlicht: (2024)
von: Yang, Yixin, et al.
Veröffentlicht: (2024)
Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought
von: Ramji, Keshav, et al.
Veröffentlicht: (2026)
von: Ramji, Keshav, et al.
Veröffentlicht: (2026)
The Landscape of Emerging AI Agent Architectures for Reasoning, Planning, and Tool Calling: A Survey
von: Masterman, Tula, et al.
Veröffentlicht: (2024)
von: Masterman, Tula, et al.
Veröffentlicht: (2024)
Fine-Tuning on Diverse Reasoning Chains Drives Within-Inference CoT Refinement in LLMs
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs
von: Zhu, Mengdan, et al.
Veröffentlicht: (2026)
von: Zhu, Mengdan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Chain-of-Thought Tokens are Computer Program Variables
von: Zhu, Fangwei, et al.
Veröffentlicht: (2025) -
Reducing Hallucinations in Entity Abstract Summarization with Facts-Template Decomposition
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024) -
Language Models Encode the Value of Numbers Linearly
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024) -
RICo: Refined In-Context Contribution for Automatic Instruction-Tuning Data Selection
von: Yang, Yixin, et al.
Veröffentlicht: (2025) -
Chip-Tuning: Classify Before Language Models Say
von: Zhu, Fangwei, et al.
Veröffentlicht: (2024)