Saved in:
| Main Authors: | Dinardi, Raul Cavalcante, Yamamoto, Bruno, Costa, Anna Helena Reali, Jordao, Artur |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2510.21067 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Effective Layer Pruning Through Similarity Metric Perspective
by: Pons, Ian, et al.
Published: (2024)
by: Pons, Ian, et al.
Published: (2024)
One Period to Rule Them All: Identifying Critical Learning Periods in Deep Networks
by: Fukase, Vinicius Yuiti, et al.
Published: (2025)
by: Fukase, Vinicius Yuiti, et al.
Published: (2025)
Layer-wise LoRA fine-tuning: a similarity metric approach
by: Ogawa, Keith Ando, et al.
Published: (2026)
by: Ogawa, Keith Ando, et al.
Published: (2026)
Efficient LLMs with AMP: Attention Heads and MLP Pruning
by: Mugnaini, Leandro Giusti, et al.
Published: (2025)
by: Mugnaini, Leandro Giusti, et al.
Published: (2025)
Technical Report on Text Dataset Distillation
by: Ogawa, Keith Ando, et al.
Published: (2025)
by: Ogawa, Keith Ando, et al.
Published: (2025)
Layer Pruning with Consensus: A Triple-Win Solution
by: Mugnaini, Leandro Giusti, et al.
Published: (2024)
by: Mugnaini, Leandro Giusti, et al.
Published: (2024)
Improving Fairness in LLMs Through Testing-Time Adversaries
by: Gregio, Isabela Pereira, et al.
Published: (2025)
by: Gregio, Isabela Pereira, et al.
Published: (2025)
Compressing LLMs with MoP: Mixture of Pruners
by: Yamamoto, Bruno Lopes, et al.
Published: (2026)
by: Yamamoto, Bruno Lopes, et al.
Published: (2026)
Realistic Market Impact Modeling for Reinforcement Learning Trading Environments
by: Abbade, Lucas Riera, et al.
Published: (2026)
by: Abbade, Lucas Riera, et al.
Published: (2026)
Comparing Normalization Methods for Portfolio Optimization with Reinforcement Learning
by: Costa, Caio de Souza Barbosa, et al.
Published: (2025)
by: Costa, Caio de Souza Barbosa, et al.
Published: (2025)
Learning to Stop Overthinking at Test Time
by: Bao, Hieu Tran, et al.
Published: (2025)
by: Bao, Hieu Tran, et al.
Published: (2025)
Pruning Everything, Everywhere, All at Once
by: Nascimento, Gustavo Henrique do, et al.
Published: (2025)
by: Nascimento, Gustavo Henrique do, et al.
Published: (2025)
When Layers Play the Lottery, all Tickets Win at Initialization
by: Jordao, Artur, et al.
Published: (2023)
by: Jordao, Artur, et al.
Published: (2023)
Don't "Overthink" Passage Reranking: Is Reasoning Truly Necessary?
by: Jedidi, Nour, et al.
Published: (2025)
by: Jedidi, Nour, et al.
Published: (2025)
Mitigating Overthinking in Large Reasoning Models via Manifold Steering
by: Huang, Yao, et al.
Published: (2025)
by: Huang, Yao, et al.
Published: (2025)
Parallel Test-Time Scaling for Latent Reasoning Models
by: You, Runyang, et al.
Published: (2025)
by: You, Runyang, et al.
Published: (2025)
Mitigating Overthinking in Large Reasoning Models via Difficulty-aware Reinforcement Learning
by: Wan, Qian, et al.
Published: (2026)
by: Wan, Qian, et al.
Published: (2026)
No Argument Left Behind: Overlapping Chunks for Faster Processing of Arbitrarily Long Legal Texts
by: Fama, Israel, et al.
Published: (2024)
by: Fama, Israel, et al.
Published: (2024)
PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning
by: Hu, Jingcheng, et al.
Published: (2026)
by: Hu, Jingcheng, et al.
Published: (2026)
Missing Premise exacerbates Overthinking: Are Reasoning Models losing Critical Thinking Skill?
by: Fan, Chenrui, et al.
Published: (2025)
by: Fan, Chenrui, et al.
Published: (2025)
SSR: Speculative Parallel Scaling Reasoning in Test-time
by: Chu, Yuanlin, et al.
Published: (2025)
by: Chu, Yuanlin, et al.
Published: (2025)
From Random to Informed Data Selection: A Diversity-Based Approach to Optimize Human Annotation and Few-Shot Learning
by: Alcoforado, Alexandre, et al.
Published: (2024)
by: Alcoforado, Alexandre, et al.
Published: (2024)
The Virtues of Pessimism in Inverse Reinforcement Learning
by: Wu, David, et al.
Published: (2024)
by: Wu, David, et al.
Published: (2024)
Right for the Right Reasons: Avoiding Reasoning Shortcuts via Prototypical Neurosymbolic AI
by: Andolfi, Luca, et al.
Published: (2025)
by: Andolfi, Luca, et al.
Published: (2025)
Implicit Reasoning in Deep Time Series Forecasting
by: Potosnak, Willa, et al.
Published: (2024)
by: Potosnak, Willa, et al.
Published: (2024)
The Virtue of Sparsity in Complexity
by: Afsharhajari, Nima, et al.
Published: (2026)
by: Afsharhajari, Nima, et al.
Published: (2026)
Investigating Compositional Reasoning in Time Series Foundation Models
by: Potosnak, Willa, et al.
Published: (2025)
by: Potosnak, Willa, et al.
Published: (2025)
Mitigating Temporal Blindness in Kubernetes Autoscaling: An Attention-Double-LSTM Framework
by: Shaikh, Faraz, et al.
Published: (2026)
by: Shaikh, Faraz, et al.
Published: (2026)
POT: Inducing Overthinking in LLMs via Black-Box Iterative Optimization
by: Li, Xinyu, et al.
Published: (2025)
by: Li, Xinyu, et al.
Published: (2025)
ROM: Real-time Overthinking Mitigation via Streaming Detection and Intervention
by: Wang, Xinyan, et al.
Published: (2026)
by: Wang, Xinyan, et al.
Published: (2026)
Overthinking the Truth: Understanding how Language Models Process False Demonstrations
by: Halawi, Danny, et al.
Published: (2023)
by: Halawi, Danny, et al.
Published: (2023)
The Cost of Avoiding Backpropagation
by: Panchal, Kunjal, et al.
Published: (2025)
by: Panchal, Kunjal, et al.
Published: (2025)
Bias as a Virtue: Rethinking Generalization under Distribution Shifts
by: Chen, Ruixuan, et al.
Published: (2025)
by: Chen, Ruixuan, et al.
Published: (2025)
Explore Briefly, Then Decide: Mitigating LLM Overthinking via Cumulative Entropy Regulation
by: Bin, Yi, et al.
Published: (2025)
by: Bin, Yi, et al.
Published: (2025)
$\nabla$-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space
by: Wang, Peihao, et al.
Published: (2026)
by: Wang, Peihao, et al.
Published: (2026)
Limits and Gains of Test-Time Scaling in Vision-Language Reasoning
by: Ahmadpour, Mohammadjavad, et al.
Published: (2025)
by: Ahmadpour, Mohammadjavad, et al.
Published: (2025)
Reward Model Generalization for Compute-Aware Test-Time Reasoning
by: Song, Zeen, et al.
Published: (2025)
by: Song, Zeen, et al.
Published: (2025)
Discrepancies are Virtue: Weak-to-Strong Generalization through Lens of Intrinsic Dimension
by: Dong, Yijun, et al.
Published: (2025)
by: Dong, Yijun, et al.
Published: (2025)
Ranking Reasoning LLMs under Test-Time Scaling
by: Hariri, Mohsen, et al.
Published: (2026)
by: Hariri, Mohsen, et al.
Published: (2026)
FastTTS: Accelerating Test-Time Scaling for Edge LLM Reasoning
by: Chen, Hao Mark, et al.
Published: (2025)
by: Chen, Hao Mark, et al.
Published: (2025)
Similar Items
-
Effective Layer Pruning Through Similarity Metric Perspective
by: Pons, Ian, et al.
Published: (2024) -
One Period to Rule Them All: Identifying Critical Learning Periods in Deep Networks
by: Fukase, Vinicius Yuiti, et al.
Published: (2025) -
Layer-wise LoRA fine-tuning: a similarity metric approach
by: Ogawa, Keith Ando, et al.
Published: (2026) -
Efficient LLMs with AMP: Attention Heads and MLP Pruning
by: Mugnaini, Leandro Giusti, et al.
Published: (2025) -
Technical Report on Text Dataset Distillation
by: Ogawa, Keith Ando, et al.
Published: (2025)