Scaling Graph Chain-of-Thought Reasoning: A Multi-Agent Framework with Efficient LLM Serving
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huan, Chengying, Meng, Ziheng, Liu, Yongchao, Yang, Zhengyi, Zhu, Yun, Yun, Yue, Li, Shipeng, Gu, Rong, Wu, Xiabao, Zhang, Haitao, Hong, Chuntao, Ma, Shaonan, Chen, Guihai, Tian, Chen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Distributed Graph Neural Network Inference With Just-In-Time Compilation For Industry-Scale Graphs
von: Wu, Xiabao, et al.
Veröffentlicht: (2025)
von: Wu, Xiabao, et al.
Veröffentlicht: (2025)
OrchANN: A Unified I/O Orchestration Framework for Skewed Out-of-Core Vector Search
von: Huan, Chengying, et al.
Veröffentlicht: (2025)
von: Huan, Chengying, et al.
Veröffentlicht: (2025)
STAR: Decode-Phase Rescheduling for LLM Inference
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
Revisiting Service Level Objectives and System Level Metrics in Large Language Model Serving
von: Wang, Zhibin, et al.
Veröffentlicht: (2024)
von: Wang, Zhibin, et al.
Veröffentlicht: (2024)
Injecting Salesperson's Dialogue Strategies in Large Language Models with Chain-of-Thought Reasoning
von: Chang, Wen-Yu, et al.
Veröffentlicht: (2024)
von: Chang, Wen-Yu, et al.
Veröffentlicht: (2024)
Enhancing Automatic Chord Recognition through LLM Chain-of-Thought Reasoning
von: Chang, Chih-Cheng, et al.
Veröffentlicht: (2025)
von: Chang, Chih-Cheng, et al.
Veröffentlicht: (2025)
HyperKAN: Hypergraph Representation Learning with Kolmogorov-Arnold Networks
von: Fang, Xiangfei, et al.
Veröffentlicht: (2025)
von: Fang, Xiangfei, et al.
Veröffentlicht: (2025)
Imitation Game for Adversarial Disillusion with Chain-of-Thought Reasoning in Generative AI
von: Chang, Ching-Chun, et al.
Veröffentlicht: (2025)
von: Chang, Ching-Chun, et al.
Veröffentlicht: (2025)
GraphGen+: Advancing Distributed Subgraph Generation and Graph Learning On Industrial Graphs
von: Jin, Yue, et al.
Veröffentlicht: (2025)
von: Jin, Yue, et al.
Veröffentlicht: (2025)
ETR: Entropy Trend Reward for Efficient Chain-of-Thought Reasoning
von: Xiong, Xuan, et al.
Veröffentlicht: (2026)
von: Xiong, Xuan, et al.
Veröffentlicht: (2026)
VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning
von: Li, Lingxiao, et al.
Veröffentlicht: (2025)
von: Li, Lingxiao, et al.
Veröffentlicht: (2025)
Graph Triple Attention Network: A Decoupled Perspective
von: Wang, Xiaotang, et al.
Veröffentlicht: (2024)
von: Wang, Xiaotang, et al.
Veröffentlicht: (2024)
SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning
von: Hu, Chenzhi, et al.
Veröffentlicht: (2026)
von: Hu, Chenzhi, et al.
Veröffentlicht: (2026)
Fast Chain-of-Thought: A Glance of Future from Parallel Decoding Leads to Answers Faster
von: Zhang, Hongxuan, et al.
Veröffentlicht: (2023)
von: Zhang, Hongxuan, et al.
Veröffentlicht: (2023)
Expanding Reasoning Potential in Foundation Model by Learning Diverse Chains of Thought Patterns
von: Zhang, Xuemiao, et al.
Veröffentlicht: (2025)
von: Zhang, Xuemiao, et al.
Veröffentlicht: (2025)
Stop Reasoning! When Multimodal LLM with Chain-of-Thought Reasoning Meets Adversarial Image
von: Wang, Zefeng, et al.
Veröffentlicht: (2024)
von: Wang, Zefeng, et al.
Veröffentlicht: (2024)
Quality-Driven Agentic Reasoning for LLM-Assisted Software Design: Questions-of-Thoughts (QoT) as a Time-Series Self-QA Chain
von: Liu, Yen-Ku, et al.
Veröffentlicht: (2026)
von: Liu, Yen-Ku, et al.
Veröffentlicht: (2026)
LLM Reasoning Is Latent, Not the Chain of Thought
von: Wang, Wenshuo
Veröffentlicht: (2026)
von: Wang, Wenshuo
Veröffentlicht: (2026)
TokenDance: Scaling Multi-Agent LLM Serving via Collective KV Cache Sharing
von: Bian, Zhuohang, et al.
Veröffentlicht: (2026)
von: Bian, Zhuohang, et al.
Veröffentlicht: (2026)
When the Chain Breaks: Interactive Diagnosis of LLM Chain-of-Thought Reasoning Errors
von: Chen, Shiwei, et al.
Veröffentlicht: (2026)
von: Chen, Shiwei, et al.
Veröffentlicht: (2026)
Rethinking and Benchmarking Large Language Models for Graph Reasoning
von: Hu, Yuwei, et al.
Veröffentlicht: (2025)
von: Hu, Yuwei, et al.
Veröffentlicht: (2025)
VG-CoT: Towards Trustworthy Visual Reasoning via Grounded Chain-of-Thought
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
von: Lim, Byeonggeuk, et al.
Veröffentlicht: (2026)
Enhancing Video-LLM Reasoning via Agent-of-Thoughts Distillation
von: Shi, Yudi, et al.
Veröffentlicht: (2024)
von: Shi, Yudi, et al.
Veröffentlicht: (2024)
EndoCoT: Scaling Endogenous Chain-of-Thought Reasoning in Diffusion Models
von: Dai, Xuanlang, et al.
Veröffentlicht: (2026)
von: Dai, Xuanlang, et al.
Veröffentlicht: (2026)
Unveiling Confirmation Bias in Chain-of-Thought Reasoning
von: Wan, Yue, et al.
Veröffentlicht: (2025)
von: Wan, Yue, et al.
Veröffentlicht: (2025)
WebCoT: Enhancing Web Agent Reasoning by Reconstructing Chain-of-Thought in Reflection, Branching, and Rollback
von: Hu, Minda, et al.
Veröffentlicht: (2025)
von: Hu, Minda, et al.
Veröffentlicht: (2025)
Echo: Efficient Co-Scheduling of Hybrid Online-Offline Tasks for Large Language Model Serving
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding
von: Chen, Xi, et al.
Veröffentlicht: (2025)
von: Chen, Xi, et al.
Veröffentlicht: (2025)
NPG-Muse: Scaling Long Chain-of-Thought Reasoning with NP-Hard Graph Problems
von: Wang, Yuyao, et al.
Veröffentlicht: (2025)
von: Wang, Yuyao, et al.
Veröffentlicht: (2025)
Efficient Reasoning via Chain of Unconscious Thought
von: Gong, Ruihan, et al.
Veröffentlicht: (2025)
von: Gong, Ruihan, et al.
Veröffentlicht: (2025)
Demystifying Long Chain-of-Thought Reasoning in LLMs
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
von: Yeo, Edward, et al.
Veröffentlicht: (2025)
Reason from Future: Reverse Thought Chain Enhances LLM Reasoning
von: Xu, Yinlong, et al.
Veröffentlicht: (2025)
von: Xu, Yinlong, et al.
Veröffentlicht: (2025)
CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning
von: Chen, Shuxu, et al.
Veröffentlicht: (2026)
von: Chen, Shuxu, et al.
Veröffentlicht: (2026)
CoT-Segmenter: Enhancing OOD Detection in Dense Road Scenes via Chain-of-Thought Reasoning
von: Song, Jeonghyo, et al.
Veröffentlicht: (2025)
von: Song, Jeonghyo, et al.
Veröffentlicht: (2025)
Pause and Reflect: Conformal Aggregation for Chain-of-Thought Reasoning
von: Gu, Yu, et al.
Veröffentlicht: (2026)
von: Gu, Yu, et al.
Veröffentlicht: (2026)
Recall with Reasoning: Chain-of-Thought Distillation for Mamba's Long-Context Memory and Extrapolation
von: Ma, Junyu, et al.
Veröffentlicht: (2025)
von: Ma, Junyu, et al.
Veröffentlicht: (2025)
Enhancing Chain-of-Thought Reasoning with Critical Representation Fine-tuning
von: Huang, Chenxi, et al.
Veröffentlicht: (2025)
von: Huang, Chenxi, et al.
Veröffentlicht: (2025)
KunServe: Parameter-centric Memory Management for Efficient Memory Overloading Handling in LLM Serving
von: Cheng, Rongxin, et al.
Veröffentlicht: (2024)
von: Cheng, Rongxin, et al.
Veröffentlicht: (2024)
CoDec: Prefix-Shared Decoding Kernel for LLMs
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
von: Wang, Zhibin, et al.
Veröffentlicht: (2025)
MPIC: Position-Independent Multimodal Context Caching System for Efficient MLLM Serving
von: Zhao, Shiju, et al.
Veröffentlicht: (2025)
von: Zhao, Shiju, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Distributed Graph Neural Network Inference With Just-In-Time Compilation For Industry-Scale Graphs
von: Wu, Xiabao, et al.
Veröffentlicht: (2025) -
OrchANN: A Unified I/O Orchestration Framework for Skewed Out-of-Core Vector Search
von: Huan, Chengying, et al.
Veröffentlicht: (2025) -
STAR: Decode-Phase Rescheduling for LLM Inference
von: Wang, Zhibin, et al.
Veröffentlicht: (2025) -
Revisiting Service Level Objectives and System Level Metrics in Large Language Model Serving
von: Wang, Zhibin, et al.
Veröffentlicht: (2024) -
Injecting Salesperson's Dialogue Strategies in Large Language Models with Chain-of-Thought Reasoning
von: Chang, Wen-Yu, et al.
Veröffentlicht: (2024)