Soundness-Aware Level: A Microscopic Signature that Predicts LLM Reasoning Potential
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Xuansheng, Pan, Xiaoman, Yao, Wenlin, Chen, Jianshu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
From Language Modeling to Instruction Following: Understanding the Behavior Shift in LLMs after Instruction Tuning
von: Wu, Xuansheng, et al.
Veröffentlicht: (2023)
von: Wu, Xuansheng, et al.
Veröffentlicht: (2023)
Usable XAI: 10 Strategies Towards Exploiting Explainability in the LLM Era
von: Wu, Xuansheng, et al.
Veröffentlicht: (2024)
von: Wu, Xuansheng, et al.
Veröffentlicht: (2024)
Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment
von: Yang, Rui, et al.
Veröffentlicht: (2024)
von: Yang, Rui, et al.
Veröffentlicht: (2024)
Thrust: Adaptively Propels Large Language Models with External Knowledge
von: Zhao, Xinran, et al.
Veröffentlicht: (2023)
von: Zhao, Xinran, et al.
Veröffentlicht: (2023)
Probabilistic Soundness Guarantees in LLM Reasoning Chains
von: You, Weiqiu, et al.
Veröffentlicht: (2025)
von: You, Weiqiu, et al.
Veröffentlicht: (2025)
Fact-and-Reflection (FaR) Improves Confidence Calibration of Large Language Models
von: Zhao, Xinran, et al.
Veröffentlicht: (2024)
von: Zhao, Xinran, et al.
Veröffentlicht: (2024)
DOTS: Learning to Reason Dynamically in LLMs via Optimal Reasoning Trajectories Search
von: Yue, Murong, et al.
Veröffentlicht: (2024)
von: Yue, Murong, et al.
Veröffentlicht: (2024)
A Comparative Analysis of LLM Memorization at Statistical and Internal Levels: Cross-Model Commonalities and Model-Specific Signatures
von: Chen, Bowen, et al.
Veröffentlicht: (2026)
von: Chen, Bowen, et al.
Veröffentlicht: (2026)
Nudging the Boundaries of LLM Reasoning
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2025)
von: Chen, Justin Chih-Yao, et al.
Veröffentlicht: (2025)
Token-Budget-Aware LLM Reasoning
von: Han, Tingxu, et al.
Veröffentlicht: (2024)
von: Han, Tingxu, et al.
Veröffentlicht: (2024)
A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models
von: Shu, Dong, et al.
Veröffentlicht: (2025)
von: Shu, Dong, et al.
Veröffentlicht: (2025)
LLM Unlearning Under the Microscope: A Full-Stack View on Methods and Metrics
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
von: Fan, Chongyu, et al.
Veröffentlicht: (2025)
Beyond Input Activations: Identifying Influential Latents by Gradient Sparse Autoencoders
von: Shu, Dong, et al.
Veröffentlicht: (2025)
von: Shu, Dong, et al.
Veröffentlicht: (2025)
Uncertainty-Aware Answer Selection for Improved Reasoning in Multi-LLM Systems
von: Agrawal, Aakriti, et al.
Veröffentlicht: (2025)
von: Agrawal, Aakriti, et al.
Veröffentlicht: (2025)
Reliability-Aware Adaptive Self-Consistency for Efficient Sampling in LLM Reasoning
von: Kim, Junseok, et al.
Veröffentlicht: (2026)
von: Kim, Junseok, et al.
Veröffentlicht: (2026)
FERA: Uncertainty-Aware Federated Reasoning for Large Language Models
von: Wang, Ruhan, et al.
Veröffentlicht: (2026)
von: Wang, Ruhan, et al.
Veröffentlicht: (2026)
Interpreting and Steering LLMs with Mutual Information-based Explanations on Sparse Autoencoders
von: Wu, Xuansheng, et al.
Veröffentlicht: (2025)
von: Wu, Xuansheng, et al.
Veröffentlicht: (2025)
SAUP: Situation Awareness Uncertainty Propagation on LLM Agent
von: Zhao, Qiwei, et al.
Veröffentlicht: (2024)
von: Zhao, Qiwei, et al.
Veröffentlicht: (2024)
Your LLM Knows the Future: Uncovering Its Multi-Token Prediction Potential
von: Samragh, Mohammad, et al.
Veröffentlicht: (2025)
von: Samragh, Mohammad, et al.
Veröffentlicht: (2025)
Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
von: Wang, Yubo, et al.
Veröffentlicht: (2025)
Stabilizing Efficient Reasoning with Step-Level Advantage Selection
von: Wang, Han, et al.
Veröffentlicht: (2026)
von: Wang, Han, et al.
Veröffentlicht: (2026)
Retrieval-enhanced Knowledge Editing in Language Models for Multi-Hop Question Answering
von: Shi, Yucheng, et al.
Veröffentlicht: (2024)
von: Shi, Yucheng, et al.
Veröffentlicht: (2024)
The Surprising Effectiveness of Negative Reinforcement in LLM Reasoning
von: Zhu, Xinyu, et al.
Veröffentlicht: (2025)
von: Zhu, Xinyu, et al.
Veröffentlicht: (2025)
Thinking with Knowledge Graphs: Enhancing LLM Reasoning Through Structured Data
von: Wu, Xue, et al.
Veröffentlicht: (2024)
von: Wu, Xue, et al.
Veröffentlicht: (2024)
Knowing When to Ask: Segment-Level Credit Assignment for LLM Tool Use
von: Kumar, Abhijit, et al.
Veröffentlicht: (2026)
von: Kumar, Abhijit, et al.
Veröffentlicht: (2026)
Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2025)
von: Cheng, Zhoujun, et al.
Veröffentlicht: (2025)
Efficiently Scaling LLM Reasoning with Certaindex
von: Fu, Yichao, et al.
Veröffentlicht: (2024)
von: Fu, Yichao, et al.
Veröffentlicht: (2024)
Small Models are LLM Knowledge Triggers on Medical Tabular Prediction
von: Yan, Jiahuan, et al.
Veröffentlicht: (2024)
von: Yan, Jiahuan, et al.
Veröffentlicht: (2024)
Struc-EMB: The Potential of Structure-Aware Encoding in Language Embeddings
von: Liu, Shikun, et al.
Veröffentlicht: (2025)
von: Liu, Shikun, et al.
Veröffentlicht: (2025)
Embracing Imperfection: Simulating Students with Diverse Cognitive Levels Using LLM-based Agents
von: Wu, Tao, et al.
Veröffentlicht: (2025)
von: Wu, Tao, et al.
Veröffentlicht: (2025)
Part I: Tricks or Traps? A Deep Dive into RL for LLM Reasoning
von: Liu, Zihe, et al.
Veröffentlicht: (2025)
von: Liu, Zihe, et al.
Veröffentlicht: (2025)
Probe and Skip: Self-Predictive Token Skipping for Efficient Long-Context LLM Inference
von: Wu, Zimeng, et al.
Veröffentlicht: (2026)
von: Wu, Zimeng, et al.
Veröffentlicht: (2026)
Self-Induced Outcome Potential: Turn-Level Credit Assignment for Agents without Verifiers
von: Hu, Senkang, et al.
Veröffentlicht: (2026)
von: Hu, Senkang, et al.
Veröffentlicht: (2026)
Recitation over Reasoning: How Cutting-Edge Language Models Can Fail on Elementary School-Level Reasoning Problems?
von: Yan, Kai, et al.
Veröffentlicht: (2025)
von: Yan, Kai, et al.
Veröffentlicht: (2025)
Improving RL Exploration for LLM Reasoning through Retrospective Replay
von: Dou, Shihan, et al.
Veröffentlicht: (2025)
von: Dou, Shihan, et al.
Veröffentlicht: (2025)
SHAPE: Stage-aware Hierarchical Advantage via Potential Estimation for LLM Reasoning
von: Ai, Zhengyang, et al.
Veröffentlicht: (2026)
von: Ai, Zhengyang, et al.
Veröffentlicht: (2026)
SwS: Self-aware Weakness-driven Problem Synthesis in Reinforcement Learning for LLM Reasoning
von: Liang, Xiao, et al.
Veröffentlicht: (2025)
von: Liang, Xiao, et al.
Veröffentlicht: (2025)
Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning
von: Noël, Valentin
Veröffentlicht: (2026)
von: Noël, Valentin
Veröffentlicht: (2026)
LSPO: Length-aware Dynamic Sampling for Policy Optimization in LLM Reasoning
von: Chen, Weizhe, et al.
Veröffentlicht: (2025)
von: Chen, Weizhe, et al.
Veröffentlicht: (2025)
Exposing Pink Slime Journalism: Linguistic Signatures and Robust Detection Against LLM-Generated Threats
von: Shahriar, Sadat, et al.
Veröffentlicht: (2025)
von: Shahriar, Sadat, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
From Language Modeling to Instruction Following: Understanding the Behavior Shift in LLMs after Instruction Tuning
von: Wu, Xuansheng, et al.
Veröffentlicht: (2023) -
Usable XAI: 10 Strategies Towards Exploiting Explainability in the LLM Era
von: Wu, Xuansheng, et al.
Veröffentlicht: (2024) -
Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment
von: Yang, Rui, et al.
Veröffentlicht: (2024) -
Thrust: Adaptively Propels Large Language Models with External Knowledge
von: Zhao, Xinran, et al.
Veröffentlicht: (2023) -
Probabilistic Soundness Guarantees in LLM Reasoning Chains
von: You, Weiqiu, et al.
Veröffentlicht: (2025)