PLDR-LLMs Learn A Generalizable Tensor Operator That Can Replace Its Own Deep Neural Net At Inference
Fuente:
arXiv
Guardado en:
| Autor principal: | Gokden, Burc |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PLDR-LLMs Reason At Self-Organized Criticality
por: Gokden, Burc
Publicado: (2026)
por: Gokden, Burc
Publicado: (2026)
PLDR-LLM: Large Language Model from Power Law Decoder Representations
por: Gokden, Burc
Publicado: (2024)
por: Gokden, Burc
Publicado: (2024)
Language Models Can Predict Their Own Behavior
por: Ashok, Dhananjay, et al.
Publicado: (2025)
por: Ashok, Dhananjay, et al.
Publicado: (2025)
Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States
por: Choi, Yunho, et al.
Publicado: (2026)
por: Choi, Yunho, et al.
Publicado: (2026)
Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation
por: Manvi, Rohin, et al.
Publicado: (2024)
por: Manvi, Rohin, et al.
Publicado: (2024)
To Each (Textual Sequence) Its Own: Improving Memorized-Data Unlearning in Large Language Models
por: Barbulescu, George-Octavian, et al.
Publicado: (2024)
por: Barbulescu, George-Octavian, et al.
Publicado: (2024)
Small LLMs Do Not Learn a Generalizable Theory of Mind via Reinforcement Learning
por: Sarangi, Sneheel, et al.
Publicado: (2025)
por: Sarangi, Sneheel, et al.
Publicado: (2025)
Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies
por: Mittal, Avni
Publicado: (2026)
por: Mittal, Avni
Publicado: (2026)
LLMs Don't Know Their Own Decision Boundaries: The Unreliability of Self-Generated Counterfactual Explanations
por: Mayne, Harry, et al.
Publicado: (2025)
por: Mayne, Harry, et al.
Publicado: (2025)
Communication Compression for Tensor Parallel LLM Inference
por: Hansen-Palmus, Jan, et al.
Publicado: (2024)
por: Hansen-Palmus, Jan, et al.
Publicado: (2024)
ComplexityNet: Increasing LLM Inference Efficiency by Learning Task Complexity
por: Bae, Henry, et al.
Publicado: (2023)
por: Bae, Henry, et al.
Publicado: (2023)
Can Large Language Models Detect Methodological Flaws? Evidence from Gesture Recognition for UAV-Based Rescue Operation Based on Deep Learning
por: Varga, Domonkos
Publicado: (2026)
por: Varga, Domonkos
Publicado: (2026)
Training Language Models to Explain Their Own Computations
por: Li, Belinda Z., et al.
Publicado: (2025)
por: Li, Belinda Z., et al.
Publicado: (2025)
Can LLMs Follow Simple Rules?
por: Mu, Norman, et al.
Publicado: (2023)
por: Mu, Norman, et al.
Publicado: (2023)
The Impact of Inference Acceleration on Bias of LLMs
por: Kirsten, Elisabeth, et al.
Publicado: (2024)
por: Kirsten, Elisabeth, et al.
Publicado: (2024)
The Remarkable Robustness of LLMs: Stages of Inference?
por: Lad, Vedang, et al.
Publicado: (2024)
por: Lad, Vedang, et al.
Publicado: (2024)
Can LLMs Help Uncover Insights about LLMs? A Large-Scale, Evolving Literature Analysis of Frontier LLMs
por: Park, Jungsoo, et al.
Publicado: (2025)
por: Park, Jungsoo, et al.
Publicado: (2025)
Light-IF: Endowing LLMs with Generalizable Reasoning via Preview and Self-Checking for Complex Instruction Following
por: Wang, Chenyang, et al.
Publicado: (2025)
por: Wang, Chenyang, et al.
Publicado: (2025)
Climbing the Ladder of Reasoning: What LLMs Can-and Still Can't-Solve after SFT?
por: Sun, Yiyou, et al.
Publicado: (2025)
por: Sun, Yiyou, et al.
Publicado: (2025)
Not All Layers of LLMs Are Necessary During Inference
por: Fan, Siqi, et al.
Publicado: (2024)
por: Fan, Siqi, et al.
Publicado: (2024)
Can GRPO Help LLMs Transcend Their Pretraining Origin?
por: Ni, Kangqi, et al.
Publicado: (2025)
por: Ni, Kangqi, et al.
Publicado: (2025)
Can LLMs Convert Graphs to Text-Attributed Graphs?
por: Wang, Zehong, et al.
Publicado: (2024)
por: Wang, Zehong, et al.
Publicado: (2024)
Can Post-Training Transform LLMs into Causal Reasoners?
por: Chen, Junqi, et al.
Publicado: (2026)
por: Chen, Junqi, et al.
Publicado: (2026)
Can LLMs Reliably Simulate Human Learner Actions? A Simulation Authoring Framework for Open-Ended Learning Environments
por: Mannekote, Amogh, et al.
Publicado: (2024)
por: Mannekote, Amogh, et al.
Publicado: (2024)
Enough Coin Flips Can Make LLMs Act Bayesian
por: Gupta, Ritwik, et al.
Publicado: (2025)
por: Gupta, Ritwik, et al.
Publicado: (2025)
SelfReflect: Can LLMs Communicate Their Internal Answer Distribution?
por: Kirchhof, Michael, et al.
Publicado: (2025)
por: Kirchhof, Michael, et al.
Publicado: (2025)
Can LLMs Speak For Diverse People? Tuning LLMs via Debate to Generate Controllable Controversial Statements
por: Li, Ming, et al.
Publicado: (2024)
por: Li, Ming, et al.
Publicado: (2024)
DLO: Dynamic Layer Operation for Efficient Vertical Scaling of LLMs
por: Tan, Zhen, et al.
Publicado: (2024)
por: Tan, Zhen, et al.
Publicado: (2024)
GLASS: Global-Local Aggregation for Inference-time Sparsification of LLMs
por: Sattarifard, Amirmohsen, et al.
Publicado: (2025)
por: Sattarifard, Amirmohsen, et al.
Publicado: (2025)
Nudging: Inference-time Alignment of LLMs via Guided Decoding
por: Fei, Yu, et al.
Publicado: (2024)
por: Fei, Yu, et al.
Publicado: (2024)
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
por: DeepSeek-AI, et al.
Publicado: (2025)
por: DeepSeek-AI, et al.
Publicado: (2025)
SafetyNet: Detecting Harmful Outputs in LLMs by Modeling and Monitoring Deceptive Behaviors
por: Chaudhary, Maheep, et al.
Publicado: (2025)
por: Chaudhary, Maheep, et al.
Publicado: (2025)
The Alignment Tax: Response Homogenization in Aligned LLMs and Its Implications for Uncertainty Estimation
por: Liu, Mingyi
Publicado: (2026)
por: Liu, Mingyi
Publicado: (2026)
Limited Generalizability in Argument Mining: State-Of-The-Art Models Learn Datasets, Not Arguments
por: Feger, Marc, et al.
Publicado: (2025)
por: Feger, Marc, et al.
Publicado: (2025)
GRLO: Towards Generalizable Reinforcement Learning in Open-Ended Environments from Zero
por: Yin, Shangjian, et al.
Publicado: (2026)
por: Yin, Shangjian, et al.
Publicado: (2026)
Agile-Quant: Activation-Guided Quantization for Faster Inference of LLMs on the Edge
por: Shen, Xuan, et al.
Publicado: (2023)
por: Shen, Xuan, et al.
Publicado: (2023)
Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps
por: Waldendorf, Jonas, et al.
Publicado: (2026)
por: Waldendorf, Jonas, et al.
Publicado: (2026)
Demystifying Hybrid Thinking: Can LLMs Truly Switch Between Think and No-Think?
por: Wang, Shouren, et al.
Publicado: (2025)
por: Wang, Shouren, et al.
Publicado: (2025)
Can Stories Help LLMs Reason? Curating Information Space Through Narrative
por: Javadi, Vahid Sadiri, et al.
Publicado: (2024)
por: Javadi, Vahid Sadiri, et al.
Publicado: (2024)
RLHF Can Speak Many Languages: Unlocking Multilingual Preference Optimization for LLMs
por: Dang, John, et al.
Publicado: (2024)
por: Dang, John, et al.
Publicado: (2024)
Ejemplares similares
-
PLDR-LLMs Reason At Self-Organized Criticality
por: Gokden, Burc
Publicado: (2026) -
PLDR-LLM: Large Language Model from Power Law Decoder Representations
por: Gokden, Burc
Publicado: (2024) -
Language Models Can Predict Their Own Behavior
por: Ashok, Dhananjay, et al.
Publicado: (2025) -
Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States
por: Choi, Yunho, et al.
Publicado: (2026) -
Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation
por: Manvi, Rohin, et al.
Publicado: (2024)