A Dual-Path Architecture for Scaling Compute and Capacity in LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Frey, Markus, Shomali, Behzad, Koehler, Joachim, Ali, Mehdi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Loops and Memory in Transformers: Think Harder or Know More?
von: Frey, Markus, et al.
Veröffentlicht: (2026)
von: Frey, Markus, et al.
Veröffentlicht: (2026)
Is continuous CoT better suited for multi-lingual reasoning?
von: Bashir, Ali Hamza, et al.
Veröffentlicht: (2026)
von: Bashir, Ali Hamza, et al.
Veröffentlicht: (2026)
Can LLMs Take Retrieved Information with a Grain of Salt?
von: Shayegh, Behzad, et al.
Veröffentlicht: (2026)
von: Shayegh, Behzad, et al.
Veröffentlicht: (2026)
Single- vs. Dual-Prompt Dialogue Generation with LLMs for Job Interviews in Human Resources
von: De Baer, Joachim, et al.
Veröffentlicht: (2025)
von: De Baer, Joachim, et al.
Veröffentlicht: (2025)
The Dual-use Dilemma in LLMs: Do Empowering Ethical Capacities Make a Degraded Utility?
von: Zhang, Yiyi, et al.
Veröffentlicht: (2025)
von: Zhang, Yiyi, et al.
Veröffentlicht: (2025)
How do Scaling Laws Apply to Knowledge Graph Engineering Tasks? The Impact of Model Size on Large Language Model Performance
von: Heim, Desiree, et al.
Veröffentlicht: (2025)
von: Heim, Desiree, et al.
Veröffentlicht: (2025)
Scaling Up Summarization: Leveraging Large Language Models for Long Text Extractive Summarization
von: Hemamou, Léo, et al.
Veröffentlicht: (2024)
von: Hemamou, Léo, et al.
Veröffentlicht: (2024)
A Dual-Axis Taxonomy of Knowledge Editing for LLMs: From Mechanisms to Functions
von: Salehoof, Amir Mohammad, et al.
Veröffentlicht: (2025)
von: Salehoof, Amir Mohammad, et al.
Veröffentlicht: (2025)
KARRIEREWEGE: A Large Scale Career Path Prediction Dataset
von: Senger, Elena, et al.
Veröffentlicht: (2024)
von: Senger, Elena, et al.
Veröffentlicht: (2024)
Efficient Latent Semantic Clustering for Scaling Test-Time Computation of LLMs
von: Lee, Sungjae, et al.
Veröffentlicht: (2025)
von: Lee, Sungjae, et al.
Veröffentlicht: (2025)
The Straight and Narrow: Do LLMs Possess an Internal Moral Path?
von: Hu, Luoming, et al.
Veröffentlicht: (2026)
von: Hu, Luoming, et al.
Veröffentlicht: (2026)
QuestA: Expanding Reasoning Capacity in LLMs via Question Augmentation
von: Li, Jiazheng, et al.
Veröffentlicht: (2025)
von: Li, Jiazheng, et al.
Veröffentlicht: (2025)
The Order Effect: Investigating Prompt Sensitivity to Input Order in LLMs
von: Guan, Bryan, et al.
Veröffentlicht: (2025)
von: Guan, Bryan, et al.
Veröffentlicht: (2025)
DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding
von: Zhang, Jiebin, et al.
Veröffentlicht: (2026)
von: Zhang, Jiebin, et al.
Veröffentlicht: (2026)
Relative Scaling Laws for LLMs
von: Held, William, et al.
Veröffentlicht: (2025)
von: Held, William, et al.
Veröffentlicht: (2025)
MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation
von: Singh, Jyotika, et al.
Veröffentlicht: (2026)
von: Singh, Jyotika, et al.
Veröffentlicht: (2026)
A Dual-Layered Evaluation of Geopolitical and Cultural Bias in LLMs
von: Kim, Sean, et al.
Veröffentlicht: (2025)
von: Kim, Sean, et al.
Veröffentlicht: (2025)
LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling
von: Zheng, Tong, et al.
Veröffentlicht: (2026)
von: Zheng, Tong, et al.
Veröffentlicht: (2026)
On the importance of Data Scale in Pretraining Arabic Language Models
von: Ghaddar, Abbas, et al.
Veröffentlicht: (2024)
von: Ghaddar, Abbas, et al.
Veröffentlicht: (2024)
A Comparative analysis of Layer-wise Representational Capacity in AR and Diffusion LLMs
von: Goel, Raghavv, et al.
Veröffentlicht: (2026)
von: Goel, Raghavv, et al.
Veröffentlicht: (2026)
Decide Then Retrieve: A Training-Free Framework with Uncertainty-Guided Triggering and Dual-Path Retrieval
von: Chen, Wang, et al.
Veröffentlicht: (2026)
von: Chen, Wang, et al.
Veröffentlicht: (2026)
Mechanistic Interpretability of Large-Scale Counting in LLMs through a System-2 Strategy
von: Hasani, Hosein, et al.
Veröffentlicht: (2026)
von: Hasani, Hosein, et al.
Veröffentlicht: (2026)
Short-Path Prompting in LLMs: Analyzing Reasoning Instability and Solutions for Robust Performance
von: Tang, Zuoli, et al.
Veröffentlicht: (2025)
von: Tang, Zuoli, et al.
Veröffentlicht: (2025)
Path-Consistency with Prefix Enhancement for Efficient Inference in LLMs
von: Zhu, Jiace, et al.
Veröffentlicht: (2024)
von: Zhu, Jiace, et al.
Veröffentlicht: (2024)
Structured Token Retention and Computational Memory Paths in Large Language Models
von: Delena, Jonathan, et al.
Veröffentlicht: (2025)
von: Delena, Jonathan, et al.
Veröffentlicht: (2025)
Architecture, Not Scale: Circuit Localization in Large Language Models
von: Venkatesh, Sohan
Veröffentlicht: (2026)
von: Venkatesh, Sohan
Veröffentlicht: (2026)
When Life Gives You Samples: The Benefits of Scaling up Inference Compute for Multilingual LLMs
von: Khairi, Ammar, et al.
Veröffentlicht: (2025)
von: Khairi, Ammar, et al.
Veröffentlicht: (2025)
Scaling Efficient LLMs
von: Kausik, B. N.
Veröffentlicht: (2024)
von: Kausik, B. N.
Veröffentlicht: (2024)
PoultryLeX-Net: Domain-Adaptive Dual-Stream Transformer Architecture for Large-Scale Poultry Stakeholder Modeling
von: Afrifa, Stephen, et al.
Veröffentlicht: (2026)
von: Afrifa, Stephen, et al.
Veröffentlicht: (2026)
LLMs Beyond English: Scaling the Multilingual Capability of LLMs with Cross-Lingual Feedback
von: Lai, Wen, et al.
Veröffentlicht: (2024)
von: Lai, Wen, et al.
Veröffentlicht: (2024)
Using LLMs to label medical papers according to the CIViC evidence model
von: Hisch, Markus, et al.
Veröffentlicht: (2024)
von: Hisch, Markus, et al.
Veröffentlicht: (2024)
Fact-checking with Generative AI: A Systematic Cross-Topic Examination of LLMs Capacity to Detect Veracity of Political Information
von: Kuznetsova, Elizaveta, et al.
Veröffentlicht: (2025)
von: Kuznetsova, Elizaveta, et al.
Veröffentlicht: (2025)
PertEval: Unveiling Real Knowledge Capacity of LLMs with Knowledge-Invariant Perturbations
von: Li, Jiatong, et al.
Veröffentlicht: (2024)
von: Li, Jiatong, et al.
Veröffentlicht: (2024)
Gradient-Controlled Decoding: A Safety Guardrail for LLMs with Dual-Anchor Steering
von: Chiniya, Purva, et al.
Veröffentlicht: (2026)
von: Chiniya, Purva, et al.
Veröffentlicht: (2026)
Measurement in the Age of LLMs: An Application to Ideological Scaling
von: O'Hagan, Sean, et al.
Veröffentlicht: (2023)
von: O'Hagan, Sean, et al.
Veröffentlicht: (2023)
Expanding Computation Spaces of LLMs at Inference Time
von: Jang, Yoonna, et al.
Veröffentlicht: (2025)
von: Jang, Yoonna, et al.
Veröffentlicht: (2025)
LLMs as Method Actors: A Model for Prompt Engineering and Architecture
von: Doyle, Colin
Veröffentlicht: (2024)
von: Doyle, Colin
Veröffentlicht: (2024)
Verdict: A Library for Scaling Judge-Time Compute
von: Kalra, Nimit, et al.
Veröffentlicht: (2025)
von: Kalra, Nimit, et al.
Veröffentlicht: (2025)
CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
von: Liu, Cheng, et al.
Veröffentlicht: (2025)
MBR and QE Finetuning: Training-time Distillation of the Best and Most Expensive Decoding Methods
von: Finkelstein, Mara, et al.
Veröffentlicht: (2023)
von: Finkelstein, Mara, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Adaptive Loops and Memory in Transformers: Think Harder or Know More?
von: Frey, Markus, et al.
Veröffentlicht: (2026) -
Is continuous CoT better suited for multi-lingual reasoning?
von: Bashir, Ali Hamza, et al.
Veröffentlicht: (2026) -
Can LLMs Take Retrieved Information with a Grain of Salt?
von: Shayegh, Behzad, et al.
Veröffentlicht: (2026) -
Single- vs. Dual-Prompt Dialogue Generation with LLMs for Job Interviews in Human Resources
von: De Baer, Joachim, et al.
Veröffentlicht: (2025) -
The Dual-use Dilemma in LLMs: Do Empowering Ethical Capacities Make a Degraded Utility?
von: Zhang, Yiyi, et al.
Veröffentlicht: (2025)