CRISP: Compressing Redundancy in Chain-of-Thought via Intrinsic Saliency Pruning
Fuente:
arXiv
Salvato in:
| Autori principali: | Lan, Yangsong, Dai, Hongliang, Li, Piji |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
5W1H Extraction With Large Language Models
di: Cao, Yang, et al.
Pubblicazione: (2024)
di: Cao, Yang, et al.
Pubblicazione: (2024)
Generating Diverse Training Samples for Relation Extraction with Large Language Models
di: Li, Zexuan, et al.
Pubblicazione: (2025)
di: Li, Zexuan, et al.
Pubblicazione: (2025)
M-BRe: Discovering Training Samples for Relation Extraction from Unlabeled Texts with Large Language Models
di: Li, Zexuan, et al.
Pubblicazione: (2025)
di: Li, Zexuan, et al.
Pubblicazione: (2025)
Graph-Based Chain-of-Thought Pruning for Reducing Redundant Reflections in Reasoning LLMs
di: Yuan, Hongyuan, et al.
Pubblicazione: (2026)
di: Yuan, Hongyuan, et al.
Pubblicazione: (2026)
Characteristic AI Agents via Large Language Models
di: Wang, Xi, et al.
Pubblicazione: (2024)
di: Wang, Xi, et al.
Pubblicazione: (2024)
An Empirical Investigation of Domain Adaptation Ability for Chinese Spelling Check Models
di: Wang, Xi, et al.
Pubblicazione: (2024)
di: Wang, Xi, et al.
Pubblicazione: (2024)
Upfront Chain-of-Thought: A Cooperative Framework for Chain-of-Thought Compression
di: Li, Chengzhengxu, et al.
Pubblicazione: (2025)
di: Li, Chengzhengxu, et al.
Pubblicazione: (2025)
ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning
di: Hou, Bairu, et al.
Pubblicazione: (2025)
di: Hou, Bairu, et al.
Pubblicazione: (2025)
Concise and Sufficient Sub-Sentence Citations for Retrieval-Augmented Generation
di: Chen, Guo, et al.
Pubblicazione: (2025)
di: Chen, Guo, et al.
Pubblicazione: (2025)
R1-Compress: Long Chain-of-Thought Compression via Chunk Compression and Search
di: Wang, Yibo, et al.
Pubblicazione: (2025)
di: Wang, Yibo, et al.
Pubblicazione: (2025)
PPC-GPT: Federated Task-Specific Compression of Large Language Models via Pruning and Chain-of-Thought Distillation
di: Fan, Tao, et al.
Pubblicazione: (2025)
di: Fan, Tao, et al.
Pubblicazione: (2025)
CODI: Compressing Chain-of-Thought into Continuous Space via Self-Distillation
di: Shen, Zhenyi, et al.
Pubblicazione: (2025)
di: Shen, Zhenyi, et al.
Pubblicazione: (2025)
TokenSkip: Controllable Chain-of-Thought Compression in LLMs
di: Xia, Heming, et al.
Pubblicazione: (2025)
di: Xia, Heming, et al.
Pubblicazione: (2025)
Self-Compression of Chain-of-Thought via Multi-Agent Reinforcement Learning
di: Chen, Yiqun, et al.
Pubblicazione: (2026)
di: Chen, Yiqun, et al.
Pubblicazione: (2026)
CtrlCoT: Dual-Granularity Chain-of-Thought Compression for Controllable Reasoning
di: Fan, Zhenxuan, et al.
Pubblicazione: (2026)
di: Fan, Zhenxuan, et al.
Pubblicazione: (2026)
MathCanvas: Intrinsic Visual Chain-of-Thought for Multimodal Mathematical Reasoning
di: Shi, Weikang, et al.
Pubblicazione: (2025)
di: Shi, Weikang, et al.
Pubblicazione: (2025)
MEMoE: Enhancing Model Editing with Mixture of Experts Adaptors
di: Wang, Renzhi, et al.
Pubblicazione: (2024)
di: Wang, Renzhi, et al.
Pubblicazione: (2024)
A Systematic Evaluation of Large Language Models for Natural Language Generation Tasks
di: Ni, Xuanfan, et al.
Pubblicazione: (2024)
di: Ni, Xuanfan, et al.
Pubblicazione: (2024)
LEMoE: Advanced Mixture of Experts Adaptor for Lifelong Model Editing of Large Language Models
di: Wang, Renzhi, et al.
Pubblicazione: (2024)
di: Wang, Renzhi, et al.
Pubblicazione: (2024)
Semantic are Beacons: A Semantic Perspective for Unveiling Parameter-Efficient Fine-Tuning in Knowledge Learning
di: Wang, Renzhi, et al.
Pubblicazione: (2024)
di: Wang, Renzhi, et al.
Pubblicazione: (2024)
CoTJudger: A Graph-Driven Framework for Automatic Evaluation of Chain-of-Thought Efficiency and Redundancy in LRMs
di: Li, Siyi, et al.
Pubblicazione: (2026)
di: Li, Siyi, et al.
Pubblicazione: (2026)
Shorter, but Still Trustworthy? An Empirical Study of Chain-of-Thought Compression
di: Zeng, Lingjie, et al.
Pubblicazione: (2026)
di: Zeng, Lingjie, et al.
Pubblicazione: (2026)
Compressed Chain of Thought: Efficient Reasoning Through Dense Representations
di: Cheng, Jeffrey, et al.
Pubblicazione: (2024)
di: Cheng, Jeffrey, et al.
Pubblicazione: (2024)
CoT-Valve: Length-Compressible Chain-of-Thought Tuning
di: Ma, Xinyin, et al.
Pubblicazione: (2025)
di: Ma, Xinyin, et al.
Pubblicazione: (2025)
DeepPrune: Parallel Scaling without Inter-trace Redundancy
di: Tu, Shangqing, et al.
Pubblicazione: (2025)
di: Tu, Shangqing, et al.
Pubblicazione: (2025)
Temporal Guidance for Large Language Models
di: Zheng, Hong-Kai, et al.
Pubblicazione: (2026)
di: Zheng, Hong-Kai, et al.
Pubblicazione: (2026)
ConMax: Confidence-Maximizing Compression for Efficient Chain-of-Thought Reasoning
di: Hu, Minda, et al.
Pubblicazione: (2026)
di: Hu, Minda, et al.
Pubblicazione: (2026)
Think Clearly: Improving Reasoning via Redundant Token Pruning
di: Choi, Daewon, et al.
Pubblicazione: (2025)
di: Choi, Daewon, et al.
Pubblicazione: (2025)
Improve Language Model and Brain Alignment via Associative Memory
di: Yin, Congchi, et al.
Pubblicazione: (2025)
di: Yin, Congchi, et al.
Pubblicazione: (2025)
Efficient Reasoning via Chain of Unconscious Thought
di: Gong, Ruihan, et al.
Pubblicazione: (2025)
di: Gong, Ruihan, et al.
Pubblicazione: (2025)
Redundancy, Isotropy, and Intrinsic Dimensionality of Prompt-based Text Embeddings
di: Tsukagoshi, Hayato, et al.
Pubblicazione: (2025)
di: Tsukagoshi, Hayato, et al.
Pubblicazione: (2025)
Saliency-driven Dynamic Token Pruning for Large Language Models
di: Tao, Yao, et al.
Pubblicazione: (2025)
di: Tao, Yao, et al.
Pubblicazione: (2025)
Faithful Logical Reasoning via Symbolic Chain-of-Thought
di: Xu, Jundong, et al.
Pubblicazione: (2024)
di: Xu, Jundong, et al.
Pubblicazione: (2024)
Cause-Aware Empathetic Response Generation via Chain-of-Thought Fine-Tuning
di: Chen, Xinhao, et al.
Pubblicazione: (2024)
di: Chen, Xinhao, et al.
Pubblicazione: (2024)
Reinforcement Learning for Chain of Thought Compression with One-Domain-to-All Generalization
di: Li, Hanyu, et al.
Pubblicazione: (2025)
di: Li, Hanyu, et al.
Pubblicazione: (2025)
Towards Threshold-Free KV Cache Pruning
di: Ni, Xuanfan, et al.
Pubblicazione: (2025)
di: Ni, Xuanfan, et al.
Pubblicazione: (2025)
Hallucination Mitigating for Medical Report Generation
di: Zhao, Ruoqing, et al.
Pubblicazione: (2026)
di: Zhao, Ruoqing, et al.
Pubblicazione: (2026)
Decoding the Echoes of Vision from fMRI: Memory Disentangling for Past Semantic Information
di: Xia, Runze, et al.
Pubblicazione: (2024)
di: Xia, Runze, et al.
Pubblicazione: (2024)
CSRP: Chain-of-Thought Reasoning for Chinese Text Correction via Reinforcement Learning with Efficiency-Aware Rewards
di: Tian, Wei, et al.
Pubblicazione: (2026)
di: Tian, Wei, et al.
Pubblicazione: (2026)
HMPO: Hybrid Median-length Policy Optimization for Chain-of-Thought Compression
di: Zheng, Minghui, et al.
Pubblicazione: (2026)
di: Zheng, Minghui, et al.
Pubblicazione: (2026)
Documenti analoghi
-
5W1H Extraction With Large Language Models
di: Cao, Yang, et al.
Pubblicazione: (2024) -
Generating Diverse Training Samples for Relation Extraction with Large Language Models
di: Li, Zexuan, et al.
Pubblicazione: (2025) -
M-BRe: Discovering Training Samples for Relation Extraction from Unlabeled Texts with Large Language Models
di: Li, Zexuan, et al.
Pubblicazione: (2025) -
Graph-Based Chain-of-Thought Pruning for Reducing Redundant Reflections in Reasoning LLMs
di: Yuan, Hongyuan, et al.
Pubblicazione: (2026) -
Characteristic AI Agents via Large Language Models
di: Wang, Xi, et al.
Pubblicazione: (2024)