DIVERSED: Relaxed Speculative Decoding via Dynamic Ensemble Verification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Ziyi, Kasa, Siva Rajesh, S, Ankith M, Kasa, Santhosh Kumar, Zou, Jiaru, Negi, Sumit, Zhang, Ruqi, Jiang, Nan, Song, Qifan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Anonymization-Enhanced Privacy Protection for Mobile GUI Agents: Available but Invisible
von: Zhao, Lepeng, et al.
Veröffentlicht: (2026)
von: Zhao, Lepeng, et al.
Veröffentlicht: (2026)
Advancing Expert Specialization for Better MoE
von: Guo, Hongcan, et al.
Veröffentlicht: (2025)
von: Guo, Hongcan, et al.
Veröffentlicht: (2025)
CLMN: Concept based Language Models via Neural Symbolic Reasoning
von: Yang, Yibo
Veröffentlicht: (2025)
von: Yang, Yibo
Veröffentlicht: (2025)
Data-driven Circuit Discovery for Interpretability of Language Models
von: Rai, Daking, et al.
Veröffentlicht: (2026)
von: Rai, Daking, et al.
Veröffentlicht: (2026)
Scalify: scale propagation for efficient low-precision LLM training
von: Balança, Paul, et al.
Veröffentlicht: (2024)
von: Balança, Paul, et al.
Veröffentlicht: (2024)
Extending $μ$P: Spectral Conditions for Feature Learning Across Optimizers
von: Gupta, Akshita, et al.
Veröffentlicht: (2026)
von: Gupta, Akshita, et al.
Veröffentlicht: (2026)
When Does Content-Based Routing Work? Representation Requirements for Selective Attention in Hybrid Sequence Models
von: Basu, Abhinaba
Veröffentlicht: (2026)
von: Basu, Abhinaba
Veröffentlicht: (2026)
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
von: Wang, Yihao, et al.
Veröffentlicht: (2026)
MAcPNN: Mutual Assisted Learning on Data Streams with Temporal Dependence
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
NOTAI.AI: Explainable Detection of Machine-Generated Text via Curvature and Feature Attribution
von: Breneur, Oleksandr Marchenko, et al.
Veröffentlicht: (2026)
von: Breneur, Oleksandr Marchenko, et al.
Veröffentlicht: (2026)
Rethinking the Multilingual Reasoning Gap with Layer Swap
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2026)
von: Lasbordes, Maxence, et al.
Veröffentlicht: (2026)
ProactBench: Beyond What The User Asked For
von: Harfi, Sepehr, et al.
Veröffentlicht: (2026)
von: Harfi, Sepehr, et al.
Veröffentlicht: (2026)
Project Synapse: A Hierarchical Multi-Agent Framework with Hybrid Memory for Autonomous Resolution of Last-Mile Delivery Disruptions
von: Yadav, Arin Gopalan, et al.
Veröffentlicht: (2026)
von: Yadav, Arin Gopalan, et al.
Veröffentlicht: (2026)
Harnessing non-adversarial robustness in large language models
von: Zhou, Qinghua, et al.
Veröffentlicht: (2026)
von: Zhou, Qinghua, et al.
Veröffentlicht: (2026)
ParaRNN: Unlocking Parallel Training of Nonlinear RNNs for Large Language Models
von: Danieli, Federico, et al.
Veröffentlicht: (2025)
von: Danieli, Federico, et al.
Veröffentlicht: (2025)
The Curious Case of In-Training Compression of State Space Models
von: Chahine, Makram, et al.
Veröffentlicht: (2025)
von: Chahine, Makram, et al.
Veröffentlicht: (2025)
NeuronSpark: A Spiking Neural Network Language Model with Selective State Space Dynamics
von: Tang, Zhengzheng
Veröffentlicht: (2026)
von: Tang, Zhengzheng
Veröffentlicht: (2026)
Thinking Machines: Mathematical Reasoning in the Age of LLMs
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
von: Asperti, Andrea, et al.
Veröffentlicht: (2025)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
von: Miliani, Martina, et al.
Veröffentlicht: (2025)
von: Miliani, Martina, et al.
Veröffentlicht: (2025)
Context Aware Lemmatization and Morphological Tagging Method in Turkish
von: Sayallar, Cagri
Veröffentlicht: (2025)
von: Sayallar, Cagri
Veröffentlicht: (2025)
MetaCheckGPT -- A Multi-task Hallucination Detector Using LLM Uncertainty and Meta-models
von: Mehta, Rahul, et al.
Veröffentlicht: (2024)
von: Mehta, Rahul, et al.
Veröffentlicht: (2024)
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
Prompting Encoder Models for Zero-Shot Classification: A Cross-Domain Study in Italian
von: Auriemma, Serena, et al.
Veröffentlicht: (2024)
von: Auriemma, Serena, et al.
Veröffentlicht: (2024)
The GPT-4o Shock Emotional Attachment to AI Models and Its Impact on Regulatory Acceptance: A Cross-Cultural Analysis of the Immediate Transition from GPT-4o to GPT-5
von: Naito, Hiroki
Veröffentlicht: (2025)
von: Naito, Hiroki
Veröffentlicht: (2025)
How Pruning Reshapes Features: Sparse Autoencoder Analysis of Weight-Pruned Language Models
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
von: Borobia, Hector, et al.
Veröffentlicht: (2026)
AIPsy-Affect: A Keyword-Free Clinical Stimulus Battery for Mechanistic Interpretability of Emotion in Language Models
von: Keeman, Michael
Veröffentlicht: (2026)
von: Keeman, Michael
Veröffentlicht: (2026)
The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth
von: Henry, James
Veröffentlicht: (2026)
von: Henry, James
Veröffentlicht: (2026)
A Practical Guide to Streaming Continual Learning
von: Cossu, Andrea, et al.
Veröffentlicht: (2026)
von: Cossu, Andrea, et al.
Veröffentlicht: (2026)
Large Language Models Report Subjective Experience Under Self-Referential Processing
von: Berg, Cameron, et al.
Veröffentlicht: (2025)
von: Berg, Cameron, et al.
Veröffentlicht: (2025)
Don't Look Back in Anger: MAGIC Net for Streaming Continual Learning with Temporal Dependence
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference
von: Mathew, Aby Mammen
Veröffentlicht: (2026)
von: Mathew, Aby Mammen
Veröffentlicht: (2026)
cPNN: Continuous Progressive Neural Networks for Evolving Streaming Time Series
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
von: Giannini, Federico, et al.
Veröffentlicht: (2026)
Mechanistic Analysis of Catastrophic Forgetting in Large Language Models During Continual Fine-tuning
von: Imanov, Olaf Yunus Laitinen
Veröffentlicht: (2026)
von: Imanov, Olaf Yunus Laitinen
Veröffentlicht: (2026)
Latent Object Permanence: Topological Phase Transitions, Free-Energy Principles, and Renormalization Group Flows in Deep Transformer Manifolds
von: Alpay, Faruk, et al.
Veröffentlicht: (2026)
von: Alpay, Faruk, et al.
Veröffentlicht: (2026)
Inference acceleration for large language models using "stairs" assisted greedy generation
von: Grigaliūnas, Domas, et al.
Veröffentlicht: (2024)
von: Grigaliūnas, Domas, et al.
Veröffentlicht: (2024)
Strategic Doctrine Language Models (sdLM): A Learning-System Framework for Doctrinal Consistency and Geopolitical Forecasting
von: Imanov, Olaf Yunus Laitinen, et al.
Veröffentlicht: (2026)
von: Imanov, Olaf Yunus Laitinen, et al.
Veröffentlicht: (2026)
Accelerating Suffix Jailbreak attacks with Prefix-Shared KV-cache
von: Wang, Xinhai, et al.
Veröffentlicht: (2026)
von: Wang, Xinhai, et al.
Veröffentlicht: (2026)
Probing for Representation Manifolds in Superposition
von: Modell, Alexander
Veröffentlicht: (2026)
von: Modell, Alexander
Veröffentlicht: (2026)
The Origins of Representation Manifolds in Large Language Models
von: Modell, Alexander, et al.
Veröffentlicht: (2025)
von: Modell, Alexander, et al.
Veröffentlicht: (2025)
HR-Agent: A Task-Oriented Dialogue (TOD) LLM Agent Tailored for HR Applications
von: Xu, Weijie, et al.
Veröffentlicht: (2024)
von: Xu, Weijie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Anonymization-Enhanced Privacy Protection for Mobile GUI Agents: Available but Invisible
von: Zhao, Lepeng, et al.
Veröffentlicht: (2026) -
Advancing Expert Specialization for Better MoE
von: Guo, Hongcan, et al.
Veröffentlicht: (2025) -
CLMN: Concept based Language Models via Neural Symbolic Reasoning
von: Yang, Yibo
Veröffentlicht: (2025) -
Data-driven Circuit Discovery for Interpretability of Language Models
von: Rai, Daking, et al.
Veröffentlicht: (2026) -
Scalify: scale propagation for efficient low-precision LLM training
von: Balança, Paul, et al.
Veröffentlicht: (2024)