Enhanced and Efficient Reasoning in Large Learning Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Valiant, Leslie G. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability
von: Arbuzov, Mikhail L., et al.
Veröffentlicht: (2026)
von: Arbuzov, Mikhail L., et al.
Veröffentlicht: (2026)
TRIM: Achieving Extreme Sparsity with Targeted Row-wise Iterative Metric-driven Pruning
von: Beck, Florentin, et al.
Veröffentlicht: (2025)
von: Beck, Florentin, et al.
Veröffentlicht: (2025)
Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers
von: Pratyush, Spandan
Veröffentlicht: (2026)
von: Pratyush, Spandan
Veröffentlicht: (2026)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
The Deterministic Horizon: Impossibility Results as Design Specifications for Trustworthy AI Systems
von: Guo, Dongxin
Veröffentlicht: (2026)
von: Guo, Dongxin
Veröffentlicht: (2026)
Time-to-Injury Forecasting in Elite Female Football: A DeepHit Survival Approach
von: Catterall, Victoria, et al.
Veröffentlicht: (2026)
von: Catterall, Victoria, et al.
Veröffentlicht: (2026)
Cognitive Load Limits in Large Language Models: Benchmarking Multi-Hop Reasoning
von: Adapala, Sai Teja Reddy
Veröffentlicht: (2025)
von: Adapala, Sai Teja Reddy
Veröffentlicht: (2025)
Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs
von: Kaiser, Daniel, et al.
Veröffentlicht: (2026)
von: Kaiser, Daniel, et al.
Veröffentlicht: (2026)
DRO-InstructZero: Distributionally Robust Prompt Optimization for Large Language Models
von: Li, Yangyang
Veröffentlicht: (2025)
von: Li, Yangyang
Veröffentlicht: (2025)
Aletheia: Quantifying Cognitive Conviction in Reasoning Models via Regularized Inverse Confusion Matrix
von: Fu, Fanzhe
Veröffentlicht: (2026)
von: Fu, Fanzhe
Veröffentlicht: (2026)
Contextual Integrity in LLMs via Reasoning and Reinforcement Learning
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
Recurrent Memory-Augmented Transformers with Chunked Attention for Long-Context Language Modeling
von: Kashyap, Ankit
Veröffentlicht: (2025)
von: Kashyap, Ankit
Veröffentlicht: (2025)
Generalizing Numerical Reasoning in Table Data through Operation Sketches and Self-Supervised Learning
von: Cho, Hanjun, et al.
Veröffentlicht: (2026)
von: Cho, Hanjun, et al.
Veröffentlicht: (2026)
Turning the TIDE: Cross-Architecture Distillation for Diffusion Large Language Models
von: Zhang, Gongbo, et al.
Veröffentlicht: (2026)
von: Zhang, Gongbo, et al.
Veröffentlicht: (2026)
Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models
von: Cui, Sasha, et al.
Veröffentlicht: (2025)
von: Cui, Sasha, et al.
Veröffentlicht: (2025)
Beyond Hallucinations: A Composite Score for Measuring Reliability in Open-Source Large Language Models
von: Salla, Rohit Kumar, et al.
Veröffentlicht: (2025)
von: Salla, Rohit Kumar, et al.
Veröffentlicht: (2025)
Counterfactual Likelihood Tests for Indirect Influence in Private Reasoning Channels
von: Lorup, Alexander Boesgaard
Veröffentlicht: (2026)
von: Lorup, Alexander Boesgaard
Veröffentlicht: (2026)
Beyond Pass@k: Breadth-Depth Metrics for Reasoning Boundaries
von: Dragoi, Marius, et al.
Veröffentlicht: (2025)
von: Dragoi, Marius, et al.
Veröffentlicht: (2025)
The Deterministic Horizon: When Extended Reasoning Fails and Tool Delegation Becomes Necessary
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
Harnessing Negative Signals: Reinforcement Distillation from Teacher Data for LLM Reasoning
von: Xu, Shuyao, et al.
Veröffentlicht: (2025)
von: Xu, Shuyao, et al.
Veröffentlicht: (2025)
Assessing Large Language Models on Islamic Legal Reasoning: Evidence from Inheritance Law Evaluation
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2025)
von: Bouchekif, Abdessalam, et al.
Veröffentlicht: (2025)
Enhancing Burmese News Classification with Kolmogorov-Arnold Network Head Fine-tuning
von: Aung, Thura, et al.
Veröffentlicht: (2025)
von: Aung, Thura, et al.
Veröffentlicht: (2025)
Grokking in the Wild: Data Augmentation for Real-World Multi-Hop Reasoning with Transformers
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
von: Abramov, Roman, et al.
Veröffentlicht: (2025)
SECURA: Sigmoid-Enhanced CUR Decomposition with Uninterrupted Retention and Low-Rank Adaptation in Large Language Models
von: Zhang, Yuxuan
Veröffentlicht: (2025)
von: Zhang, Yuxuan
Veröffentlicht: (2025)
Model Collapse as Cultural Evolution
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
von: Guo, Dongxin, et al.
Veröffentlicht: (2026)
TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning
von: Pan, Muyu, et al.
Veröffentlicht: (2026)
von: Pan, Muyu, et al.
Veröffentlicht: (2026)
Alternating Reinforcement Learning with Contextual Rubric Rewards: Beyond the Scalarization Strategy
von: Lan, Guangchen, et al.
Veröffentlicht: (2026)
von: Lan, Guangchen, et al.
Veröffentlicht: (2026)
Alif: Advancing Urdu Large Language Models via Multilingual Synthetic Data Distillation
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2025)
von: Shafique, Muhammad Ali, et al.
Veröffentlicht: (2025)
Adapting While Learning: Grounding LLMs for Scientific Problems with Intelligent Tool Usage Adaptation
von: Lyu, Bohan, et al.
Veröffentlicht: (2024)
von: Lyu, Bohan, et al.
Veröffentlicht: (2024)
Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features
von: Cho, Seonglae, et al.
Veröffentlicht: (2026)
von: Cho, Seonglae, et al.
Veröffentlicht: (2026)
Structured Prompt Optimization Meets Reinforcement Learning for Global and Local Interpretability over Complex Text
von: Zhou, Tianyang, et al.
Veröffentlicht: (2026)
von: Zhou, Tianyang, et al.
Veröffentlicht: (2026)
Prototype Transformer: Towards Language Model Architectures Interpretable by Design
von: Yordanov, Yordan, et al.
Veröffentlicht: (2026)
von: Yordanov, Yordan, et al.
Veröffentlicht: (2026)
Domain-Specific Pretraining of Language Models: A Comparative Study in the Medical Field
von: Kerner, Tobias
Veröffentlicht: (2024)
von: Kerner, Tobias
Veröffentlicht: (2024)
Intention Collapse: Intention-Level Metrics for Reasoning in Language Models
von: Vera, Patricio
Veröffentlicht: (2026)
von: Vera, Patricio
Veröffentlicht: (2026)
The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models
von: Liu, Ming
Veröffentlicht: (2026)
von: Liu, Ming
Veröffentlicht: (2026)
MaPPO: Maximum a Posteriori Preference Optimization with Prior Knowledge
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
von: Lan, Guangchen, et al.
Veröffentlicht: (2025)
GraphEval36K: Benchmarking Coding and Reasoning Capabilities of Large Language Models on Graph Datasets
von: Wu, Qiming, et al.
Veröffentlicht: (2024)
von: Wu, Qiming, et al.
Veröffentlicht: (2024)
Why Models Know But Don't Say: Chain-of-Thought Faithfulness Divergence Between Thinking Tokens and Answers in Open-Weight Reasoning Models
von: Young, Richard J.
Veröffentlicht: (2026)
von: Young, Richard J.
Veröffentlicht: (2026)
Reinforcement Learning for Scalable and Trustworthy Intelligent Systems
von: Lan, Guangchen
Veröffentlicht: (2026)
von: Lan, Guangchen
Veröffentlicht: (2026)
In-Context Fixation: When Demonstrated Labels Override Semantics in Few-Shot Classification
von: Liu, Ming
Veröffentlicht: (2026)
von: Liu, Ming
Veröffentlicht: (2026)
Ähnliche Einträge
-
The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability
von: Arbuzov, Mikhail L., et al.
Veröffentlicht: (2026) -
TRIM: Achieving Extreme Sparsity with Targeted Row-wise Iterative Metric-driven Pruning
von: Beck, Florentin, et al.
Veröffentlicht: (2025) -
Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers
von: Pratyush, Spandan
Veröffentlicht: (2026) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025) -
The Deterministic Horizon: Impossibility Results as Design Specifications for Trustworthy AI Systems
von: Guo, Dongxin
Veröffentlicht: (2026)