Rethinking LLM Advancement: Compute-Dependent and Independent Paths to Progress
Fuente:
arXiv
Saved in:
| Main Authors: | Sanderson, Jack, Foley, Teddy, Guo, Spencer, Qu, Anqi, Josephson, Henry |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Overclocking LLM Reasoning: Monitoring and Controlling Thinking Path Lengths in LLMs
by: Eisenstadt, Roy, et al.
Published: (2025)
by: Eisenstadt, Roy, et al.
Published: (2025)
ReflexGrad: Within-Episode Failure Recovery in LLM Agents via Progress-Gated Dual-Process Routing
by: Kadu, Ankush, et al.
Published: (2025)
by: Kadu, Ankush, et al.
Published: (2025)
Node-Level Uncertainty Estimation in LLM-Generated SQL
by: Hasson, Hilaf, et al.
Published: (2025)
by: Hasson, Hilaf, et al.
Published: (2025)
BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
by: Ustaomeroglu, Muhammed, et al.
Published: (2026)
Uncovering Bias Paths with LLM-guided Causal Discovery: An Active Learning and Dynamic Scoring Approach
by: Zanna, Khadija, et al.
Published: (2025)
by: Zanna, Khadija, et al.
Published: (2025)
SCULPT: Constraint-Guided Pruned MCTS that Carves Efficient Paths for Mathematical Reasoning
by: Fang, Qitong, et al.
Published: (2026)
by: Fang, Qitong, et al.
Published: (2026)
The ARC of Progress towards AGI: A Living Survey of Abstraction and Reasoning
by: Vahdati, Sahar, et al.
Published: (2026)
by: Vahdati, Sahar, et al.
Published: (2026)
FlowSteer: Towards Agents Designing Agentic Workflows via Reinforced Progressive Canvas Editing
by: Zhang, Mingda, et al.
Published: (2026)
by: Zhang, Mingda, et al.
Published: (2026)
Path-Coupled Bellman Flows for Distributional Reinforcement Learning
by: Xu, Boyang, et al.
Published: (2026)
by: Xu, Boyang, et al.
Published: (2026)
Interestingness as an Inductive Heuristic for Future Compression Progress
by: Herrmann, Vincent, et al.
Published: (2026)
by: Herrmann, Vincent, et al.
Published: (2026)
TensorOpera Router: A Multi-Model Router for Efficient LLM Inference
by: Stripelis, Dimitris, et al.
Published: (2024)
by: Stripelis, Dimitris, et al.
Published: (2024)
Brittlebench: Quantifying LLM robustness via prompt sensitivity
by: Romanou, Angelika, et al.
Published: (2026)
by: Romanou, Angelika, et al.
Published: (2026)
Anomaly Detection with Variance Stabilized Density Estimation
by: Rozner, Amit, et al.
Published: (2023)
by: Rozner, Amit, et al.
Published: (2023)
The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination
by: Yin, Chenlong, et al.
Published: (2025)
by: Yin, Chenlong, et al.
Published: (2025)
On the Generalization Gap in LLM Planning: Tests and Verifier-Reward RL
by: Belcamino, Valerio, et al.
Published: (2026)
by: Belcamino, Valerio, et al.
Published: (2026)
Geometric Evolution Maps: Extracting Stable Concept Probes from Transformer Residual Streams
by: Henry, James
Published: (2026)
by: Henry, James
Published: (2026)
Context-Dependent Affordance Computation in Vision-Language Models
by: Farzulla, Murad
Published: (2026)
by: Farzulla, Murad
Published: (2026)
FastForward Pruning: Efficient LLM Pruning via Single-Step Reinforcement Learning
by: Yuan, Xin, et al.
Published: (2025)
by: Yuan, Xin, et al.
Published: (2025)
AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites
by: Zhang, Qinshi, et al.
Published: (2026)
by: Zhang, Qinshi, et al.
Published: (2026)
MRI Brain Tumor Detection with Computer Vision
by: Krolik, Jack, et al.
Published: (2025)
by: Krolik, Jack, et al.
Published: (2025)
Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress
by: Agia, Christopher, et al.
Published: (2024)
by: Agia, Christopher, et al.
Published: (2024)
No Free Swap: Protocol-Dependent Layer Redundancy in Transformers
by: Garcia, Gabriel
Published: (2026)
by: Garcia, Gabriel
Published: (2026)
Symbolic Graph Networks for Robust PDE Discovery from Noisy Sparse Data
by: Chen, Xingyu, et al.
Published: (2026)
by: Chen, Xingyu, et al.
Published: (2026)
TiVaT: A Transformer with a Single Unified Mechanism for Capturing Asynchronous Dependencies in Multivariate Time Series Forecasting
by: Ha, Junwoo, et al.
Published: (2024)
by: Ha, Junwoo, et al.
Published: (2024)
LLM attribution analysis across different fine-tuning strategies and model scales for automated code compliance
by: Shi, Jack Wei Lun, et al.
Published: (2026)
by: Shi, Jack Wei Lun, et al.
Published: (2026)
ParalESN: Enabling parallel information processing in Reservoir Computing
by: Pinna, Matteo, et al.
Published: (2026)
by: Pinna, Matteo, et al.
Published: (2026)
Towards Independence Criterion in Machine Unlearning of Features and Labels
by: Han, Ling, et al.
Published: (2024)
by: Han, Ling, et al.
Published: (2024)
Beyond the Black Box: A Statistical Model for LLM Reasoning and Inference
by: Dalal, Siddhartha, et al.
Published: (2024)
by: Dalal, Siddhartha, et al.
Published: (2024)
Balancing Efficiency and Effectiveness: An LLM-Infused Approach for Optimized CTR Prediction
by: Zhang, Guoxiao, et al.
Published: (2024)
by: Zhang, Guoxiao, et al.
Published: (2024)
Solving the Granularity Mismatch: Hierarchical Preference Learning for Long-Horizon LLM Agents
by: Gao, Heyang, et al.
Published: (2025)
by: Gao, Heyang, et al.
Published: (2025)
Identity as Attractor: Geometric Evidence for Persistent Agent Architecture in LLM Activation Space
by: Vasilenko, Vladimir
Published: (2026)
by: Vasilenko, Vladimir
Published: (2026)
ObfusQAte: A Proposed Framework to Evaluate LLM Robustness on Obfuscated Factual Question Answering
by: Ghosh, Shubhra, et al.
Published: (2025)
by: Ghosh, Shubhra, et al.
Published: (2025)
Discovering physical laws with parallel symbolic enumeration
by: Ruan, Kai, et al.
Published: (2024)
by: Ruan, Kai, et al.
Published: (2024)
RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement
by: Lin, Junjie, et al.
Published: (2024)
by: Lin, Junjie, et al.
Published: (2024)
Achieving Scalable Robot Autonomy via neurosymbolic planning using lightweight local LLM
by: Attolino, Nicholas, et al.
Published: (2025)
by: Attolino, Nicholas, et al.
Published: (2025)
SaiVLA-0: Cerebrum--Pons--Cerebellum Tripartite Architecture for Compute-Aware Vision-Language-Action
by: Shi, Xiang, et al.
Published: (2026)
by: Shi, Xiang, et al.
Published: (2026)
Random Rule Forest (RRF): Interpretable Ensembles of LLM-Generated Questions for Predicting Startup Success
by: Griffin, Ben, et al.
Published: (2025)
by: Griffin, Ben, et al.
Published: (2025)
MESS+: Dynamically Learned Inference-Time LLM Routing in Model Zoos with Service Level Guarantees
by: Woisetschläger, Herbert, et al.
Published: (2025)
by: Woisetschläger, Herbert, et al.
Published: (2025)
ACE : Off-Policy Actor-Critic with Causality-Aware Entropy Regularization
by: Ji, Tianying, et al.
Published: (2024)
by: Ji, Tianying, et al.
Published: (2024)
Improving Fairness with Ensemble Combination: Margin-Dependent Bounds
by: Bian, Yijun
Published: (2023)
by: Bian, Yijun
Published: (2023)
Similar Items
-
Overclocking LLM Reasoning: Monitoring and Controlling Thinking Path Lengths in LLMs
by: Eisenstadt, Roy, et al.
Published: (2025) -
ReflexGrad: Within-Episode Failure Recovery in LLM Agents via Progress-Gated Dual-Process Routing
by: Kadu, Ankush, et al.
Published: (2025) -
Node-Level Uncertainty Estimation in LLM-Generated SQL
by: Hasson, Hilaf, et al.
Published: (2025) -
BLOCK-EM: Preventing Emergent Misalignment via Latent Blocking
by: Ustaomeroglu, Muhammed, et al.
Published: (2026) -
Uncovering Bias Paths with LLM-guided Causal Discovery: An Active Learning and Dynamic Scoring Approach
by: Zanna, Khadija, et al.
Published: (2025)