Evaluating Small Language Models for Front-Door Routing: A Harmonized Benchmark and Synthetic-Traffic Experiment
Fuente:
arXiv
Saved in:
| Main Authors: | Johnson, Warren, Lee, Charles |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Communication Traffic Characteristics Reveal an IoT Devices Identity
by: Chowdhury, Rajarshi Roy, et al.
Published: (2024)
by: Chowdhury, Rajarshi Roy, et al.
Published: (2024)
Semantic Caching for Improving Web Affordability
by: Akbar, Hafsa, et al.
Published: (2025)
by: Akbar, Hafsa, et al.
Published: (2025)
5G Network Automation Using Local Large Language Models and Retrieval-Augmented Generation
by: Majlesara, Ahmadreza, et al.
Published: (2025)
by: Majlesara, Ahmadreza, et al.
Published: (2025)
Adaptive Multi-Dimensional Coordinated Comprehensive Routing Scheme for IoV
by: Ren, Ruixing, et al.
Published: (2026)
by: Ren, Ruixing, et al.
Published: (2026)
Intelligent Channel Allocation for IEEE 802.11be Multi-Link Operation: When MAB Meets LLM
by: Lian, Shumin, et al.
Published: (2025)
by: Lian, Shumin, et al.
Published: (2025)
AIvailable: A Software-Defined Architecture for LLM-as-a-Service on Heterogeneous and Legacy GPUs
by: Antunes, Pedro, et al.
Published: (2025)
by: Antunes, Pedro, et al.
Published: (2025)
Prompt Compression in Production Task Orchestration: A Pre-Registered Randomized Trial
by: Johnson, Warren, et al.
Published: (2026)
by: Johnson, Warren, et al.
Published: (2026)
An Analysis of Active Learning Algorithms using Real-World Crowd-sourced Text Annotations
by: Totakura, Varun, et al.
Published: (2026)
by: Totakura, Varun, et al.
Published: (2026)
Compression Method Matters: Benchmark-Dependent Output Dynamics in LLM Prompt Compression
by: Johnson, Warren
Published: (2026)
by: Johnson, Warren
Published: (2026)
CBR -- Boosting Adaptive Classification By Retrieval of Encrypted Network Traffic with Out-of-distribution
by: Lukach, Amir, et al.
Published: (2024)
by: Lukach, Amir, et al.
Published: (2024)
Efficient Telecom Specific LLM: TSLAM-Mini with QLoRA and Digital Twin Data
by: Ethiraj, Vignesh, et al.
Published: (2025)
by: Ethiraj, Vignesh, et al.
Published: (2025)
Neural Router: Semantic Content Matching for Agentic AI
by: Lovén, Lauri, et al.
Published: (2026)
by: Lovén, Lauri, et al.
Published: (2026)
German Text Simplification: Finetuning Large Language Models with Semi-Synthetic Data
by: Klöser, Lars, et al.
Published: (2024)
by: Klöser, Lars, et al.
Published: (2024)
Permission Manifests for Web Agents
by: Marro, Samuele, et al.
Published: (2025)
by: Marro, Samuele, et al.
Published: (2025)
The Perplexity Paradox: Why Code Compresses Better Than Math in LLM Prompts
by: Johnson, Warren
Published: (2026)
by: Johnson, Warren
Published: (2026)
PL-Guard: Benchmarking Language Model Safety for Polish
by: Krasnodębska, Aleksandra, et al.
Published: (2025)
by: Krasnodębska, Aleksandra, et al.
Published: (2025)
MCP-Diag: A Deterministic, Protocol-Driven Architecture for AI-Native Network Diagnostics
by: Lodha, Devansh, et al.
Published: (2026)
by: Lodha, Devansh, et al.
Published: (2026)
Network and Systems Performance Characterization of MCP-Enabled LLM Agents
by: Ding, Zihao, et al.
Published: (2025)
by: Ding, Zihao, et al.
Published: (2025)
Efficient Aspect-Based Summarization of Climate Change Reports with Small Language Models
by: Ghinassi, Iacopo, et al.
Published: (2024)
by: Ghinassi, Iacopo, et al.
Published: (2024)
Luth: Efficient French Specialization for Small Language Models and Cross-Lingual Transfer
by: Lasbordes, Maxence, et al.
Published: (2025)
by: Lasbordes, Maxence, et al.
Published: (2025)
Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models
by: Gokdemir, Ozan, et al.
Published: (2025)
by: Gokdemir, Ozan, et al.
Published: (2025)
Enhancing the Reasoning Capabilities of Small Language Models via Solution Guidance Fine-Tuning
by: Bi, Jing, et al.
Published: (2024)
by: Bi, Jing, et al.
Published: (2024)
Synthetic Voice Data for Automatic Speech Recognition in African Languages
by: DeRenzi, Brian, et al.
Published: (2025)
by: DeRenzi, Brian, et al.
Published: (2025)
Tram-FL: Routing-based Model Training for Decentralized Federated Learning
by: Maejima, Kota, et al.
Published: (2023)
by: Maejima, Kota, et al.
Published: (2023)
A Multi-Task Benchmark for Abusive Language Detection in Low-Resource Settings
by: Gaim, Fitsum, et al.
Published: (2025)
by: Gaim, Fitsum, et al.
Published: (2025)
SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading
by: Dinh, Tu Anh, et al.
Published: (2024)
by: Dinh, Tu Anh, et al.
Published: (2024)
GroUSE: A Benchmark to Evaluate Evaluators in Grounded Question Answering
by: Muller, Sacha, et al.
Published: (2024)
by: Muller, Sacha, et al.
Published: (2024)
SynDocDis: A Metadata-Driven Framework for Generating Synthetic Physician Discussions Using Large Language Models
by: Rubinstein, Beny, et al.
Published: (2026)
by: Rubinstein, Beny, et al.
Published: (2026)
EnDive: A Cross-Dialect Benchmark for Fairness and Performance in Large Language Models
by: Gupta, Abhay, et al.
Published: (2025)
by: Gupta, Abhay, et al.
Published: (2025)
CoPE: A Small Language Model for Steerable and Scalable Content Labeling
by: Chakrabarti, Samidh, et al.
Published: (2025)
by: Chakrabarti, Samidh, et al.
Published: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
by: Collado-Montañez, Jaime, et al.
Published: (2025)
by: Collado-Montañez, Jaime, et al.
Published: (2025)
UA-Legal-Bench: A Benchmark for Evaluating Large Language Models on Ukrainian Legal Reasoning
by: Ovcharov, Volodymyr
Published: (2026)
by: Ovcharov, Volodymyr
Published: (2026)
Kastor: Fine-tuned Small Language Models for Shape-based Active Relation Extraction
by: Celian, Ringwald, et al.
Published: (2025)
by: Celian, Ringwald, et al.
Published: (2025)
D-COT: Disciplined Chain-of-Thought Learning for Efficient Reasoning in Small Language Models
by: Ubukata, Shunsuke
Published: (2026)
by: Ubukata, Shunsuke
Published: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)
by: Ashuach, Tomer, et al.
Published: (2025)
Towards Message Brokers for Generative AI: Survey, Challenges, and Opportunities
by: Saleh, Alaa, et al.
Published: (2023)
by: Saleh, Alaa, et al.
Published: (2023)
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
by: Karinshak, Elise, et al.
Published: (2024)
by: Karinshak, Elise, et al.
Published: (2024)
BOUQuET: dataset, Benchmark and Open initiative for Universal Quality Evaluation in Translation
by: The Omnilingual MT Team, et al.
Published: (2025)
by: The Omnilingual MT Team, et al.
Published: (2025)
RAID: A Shared Benchmark for Robust Evaluation of Machine-Generated Text Detectors
by: Dugan, Liam, et al.
Published: (2024)
by: Dugan, Liam, et al.
Published: (2024)
OPOR-Bench: Evaluating Large Language Models on Online Public Opinion Report Generation
by: Yu, Jinzheng, et al.
Published: (2025)
by: Yu, Jinzheng, et al.
Published: (2025)
Similar Items
-
Communication Traffic Characteristics Reveal an IoT Devices Identity
by: Chowdhury, Rajarshi Roy, et al.
Published: (2024) -
Semantic Caching for Improving Web Affordability
by: Akbar, Hafsa, et al.
Published: (2025) -
5G Network Automation Using Local Large Language Models and Retrieval-Augmented Generation
by: Majlesara, Ahmadreza, et al.
Published: (2025) -
Adaptive Multi-Dimensional Coordinated Comprehensive Routing Scheme for IoV
by: Ren, Ruixing, et al.
Published: (2026) -
Intelligent Channel Allocation for IEEE 802.11be Multi-Link Operation: When MAB Meets LLM
by: Lian, Shumin, et al.
Published: (2025)