Long-Document QA with Chain-of-Structured-Thought and Fine-Tuned SLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Liang, Zhuowen, Lin, Xiaotian, Zhang, Zhengxuan, Luo, Yuyu, Wang, Haixun, Tang, Nan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning
di: Li, Miao, et al.
Pubblicazione: (2026)
di: Li, Miao, et al.
Pubblicazione: (2026)
Fine-Tuning Small Language Models (SLMs) for Autonomous Web-based Geographical Information Systems (AWebGIS)
di: Ashani, Mahdi Nazari, et al.
Pubblicazione: (2025)
di: Ashani, Mahdi Nazari, et al.
Pubblicazione: (2025)
LEAD: Iterative Data Selection for Efficient LLM Instruction Tuning
di: Lin, Xiaotian, et al.
Pubblicazione: (2025)
di: Lin, Xiaotian, et al.
Pubblicazione: (2025)
Long Chain-of-Thought Reasoning Across Languages
di: Barua, Josh, et al.
Pubblicazione: (2025)
di: Barua, Josh, et al.
Pubblicazione: (2025)
NPG-Muse: Scaling Long Chain-of-Thought Reasoning with NP-Hard Graph Problems
di: Wang, Yuyao, et al.
Pubblicazione: (2025)
di: Wang, Yuyao, et al.
Pubblicazione: (2025)
Atom of Thoughts for Markov LLM Test-Time Scaling
di: Teng, Fengwei, et al.
Pubblicazione: (2025)
di: Teng, Fengwei, et al.
Pubblicazione: (2025)
A Structure-Agnostic Co-Tuning Framework for LLMs and SLMs in Cloud-Edge Systems
di: Liu, Yuze, et al.
Pubblicazione: (2025)
di: Liu, Yuze, et al.
Pubblicazione: (2025)
A Theoretical Understanding of Chain-of-Thought: Coherent Reasoning and Error-Aware Demonstration
di: Cui, Yingqian, et al.
Pubblicazione: (2024)
di: Cui, Yingqian, et al.
Pubblicazione: (2024)
ARES: Alternating Reinforcement Learning and Supervised Fine-Tuning for Enhanced Multi-Modal Chain-of-Thought Reasoning Through Diverse AI Feedback
di: Byun, Ju-Seung, et al.
Pubblicazione: (2024)
di: Byun, Ju-Seung, et al.
Pubblicazione: (2024)
Fine-Grained Knowledge Structuring and Retrieval for Visual Question Answering
di: Zhang, Zhengxuan, et al.
Pubblicazione: (2025)
di: Zhang, Zhengxuan, et al.
Pubblicazione: (2025)
Diffusion of Thoughts: Chain-of-Thought Reasoning in Diffusion Language Models
di: Ye, Jiacheng, et al.
Pubblicazione: (2024)
di: Ye, Jiacheng, et al.
Pubblicazione: (2024)
Optimizing Chain-of-Thought Reasoners via Gradient Variance Minimization in Rejection Sampling and RL
di: Yao, Jiarui, et al.
Pubblicazione: (2025)
di: Yao, Jiarui, et al.
Pubblicazione: (2025)
Aggregated Knowledge Model: Enhancing Domain-Specific QA with Fine-Tuned and Retrieval-Augmented Generation Models
di: Liu, Fengchen, et al.
Pubblicazione: (2024)
di: Liu, Fengchen, et al.
Pubblicazione: (2024)
Is Chain-of-Thought Really Not Explainability? Chain-of-Thought Can Be Faithful without Hint Verbalization
di: Zaman, Kerem, et al.
Pubblicazione: (2025)
di: Zaman, Kerem, et al.
Pubblicazione: (2025)
Enhancing Reasoning Capabilities in SLMs with Reward Guided Dataset Distillation
di: Padarha, Shreyansh
Pubblicazione: (2025)
di: Padarha, Shreyansh
Pubblicazione: (2025)
Reinforcement Learning for Chain of Thought Compression with One-Domain-to-All Generalization
di: Li, Hanyu, et al.
Pubblicazione: (2025)
di: Li, Hanyu, et al.
Pubblicazione: (2025)
Learning to Rank Chain-of-Thought: Using a Small Model
di: Jiang, Eric Hanchen, et al.
Pubblicazione: (2025)
di: Jiang, Eric Hanchen, et al.
Pubblicazione: (2025)
Fractured Chain-of-Thought Reasoning
di: Liao, Baohao, et al.
Pubblicazione: (2025)
di: Liao, Baohao, et al.
Pubblicazione: (2025)
HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA
di: Chen, Xinyue, et al.
Pubblicazione: (2024)
di: Chen, Xinyue, et al.
Pubblicazione: (2024)
Understanding Reasoning in Chain-of-Thought from the Hopfieldian View
di: Hu, Lijie, et al.
Pubblicazione: (2024)
di: Hu, Lijie, et al.
Pubblicazione: (2024)
Stepwise Perplexity-Guided Refinement for Efficient Chain-of-Thought Reasoning in Large Language Models
di: Cui, Yingqian, et al.
Pubblicazione: (2025)
di: Cui, Yingqian, et al.
Pubblicazione: (2025)
Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoning
di: Li, Xintong, et al.
Pubblicazione: (2026)
di: Li, Xintong, et al.
Pubblicazione: (2026)
Parameter-Efficient Fine-Tuning for Foundation Models
di: Zhang, Dan, et al.
Pubblicazione: (2025)
di: Zhang, Dan, et al.
Pubblicazione: (2025)
Demystifying Chains, Trees, and Graphs of Thoughts
di: Besta, Maciej, et al.
Pubblicazione: (2024)
di: Besta, Maciej, et al.
Pubblicazione: (2024)
A Formal Comparison Between Chain of Thought and Latent Thought
di: Xu, Kevin, et al.
Pubblicazione: (2025)
di: Xu, Kevin, et al.
Pubblicazione: (2025)
TopicTag: Automatic Annotation of NMF Topic Models Using Chain of Thought and Prompt Tuning with LLMs
di: Wanna, Selma, et al.
Pubblicazione: (2024)
di: Wanna, Selma, et al.
Pubblicazione: (2024)
DeferMem: Query-Time Evidence Distillation via Reinforcement Learning for Long-Term Memory QA
di: Yin, Jianing, et al.
Pubblicazione: (2026)
di: Yin, Jianing, et al.
Pubblicazione: (2026)
Think When You Need: Self-Adaptive Chain-of-Thought Learning
di: Yang, Junjie, et al.
Pubblicazione: (2025)
di: Yang, Junjie, et al.
Pubblicazione: (2025)
FedProxy: Federated Fine-Tuning of LLMs via Proxy SLMs and Heterogeneity-Aware Fusion
di: Fan, Tao, et al.
Pubblicazione: (2026)
di: Fan, Tao, et al.
Pubblicazione: (2026)
Chain-of-Thought Unfaithfulness as Disguised Accuracy
di: Bentham, Oliver, et al.
Pubblicazione: (2024)
di: Bentham, Oliver, et al.
Pubblicazione: (2024)
Generating Chain-of-Thoughts with a Pairwise-Comparison Approach to Searching for the Most Promising Intermediate Thought
di: Zhang, Zhen-Yu, et al.
Pubblicazione: (2024)
di: Zhang, Zhen-Yu, et al.
Pubblicazione: (2024)
Linear Chain Transformation: Expanding Optimization Dynamics for Fine-Tuning Large Language Models
di: Wang, Yulong, et al.
Pubblicazione: (2024)
di: Wang, Yulong, et al.
Pubblicazione: (2024)
DRAG: Distilling RAG for SLMs from LLMs to Transfer Knowledge and Mitigate Hallucination via Evidence and Graph-based Distillation
di: Chen, Jennifer, et al.
Pubblicazione: (2025)
di: Chen, Jennifer, et al.
Pubblicazione: (2025)
Long Exposure: Accelerating Parameter-Efficient Fine-Tuning for LLMs under Shadowy Sparsity
di: Wang, Tuowei, et al.
Pubblicazione: (2025)
di: Wang, Tuowei, et al.
Pubblicazione: (2025)
Let Me Think! A Long Chain-of-Thought Can Be Worth Exponentially Many Short Ones
di: Mirtaheri, Parsa, et al.
Pubblicazione: (2025)
di: Mirtaheri, Parsa, et al.
Pubblicazione: (2025)
Fragile Thoughts: How Large Language Models Handle Chain-of-Thought Perturbations
di: Aravindan, Ashwath Vaithinathan, et al.
Pubblicazione: (2026)
di: Aravindan, Ashwath Vaithinathan, et al.
Pubblicazione: (2026)
Continuous QA Learning with Structured Prompts
di: Zheng, Yinhe
Pubblicazione: (2022)
di: Zheng, Yinhe
Pubblicazione: (2022)
LLM Maybe LongLM: Self-Extend LLM Context Window Without Tuning
di: Jin, Hongye, et al.
Pubblicazione: (2024)
di: Jin, Hongye, et al.
Pubblicazione: (2024)
FinChain: A Symbolic Benchmark for Verifiable Chain-of-Thought Financial Reasoning
di: Xie, Zhuohan, et al.
Pubblicazione: (2025)
di: Xie, Zhuohan, et al.
Pubblicazione: (2025)
Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
di: Arcuschin, Iván, et al.
Pubblicazione: (2025)
di: Arcuschin, Iván, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning
di: Li, Miao, et al.
Pubblicazione: (2026) -
Fine-Tuning Small Language Models (SLMs) for Autonomous Web-based Geographical Information Systems (AWebGIS)
di: Ashani, Mahdi Nazari, et al.
Pubblicazione: (2025) -
LEAD: Iterative Data Selection for Efficient LLM Instruction Tuning
di: Lin, Xiaotian, et al.
Pubblicazione: (2025) -
Long Chain-of-Thought Reasoning Across Languages
di: Barua, Josh, et al.
Pubblicazione: (2025) -
NPG-Muse: Scaling Long Chain-of-Thought Reasoning with NP-Hard Graph Problems
di: Wang, Yuyao, et al.
Pubblicazione: (2025)