Streaming, Fast and Slow: Cognitive Load-Aware Streaming for Efficient LLM Serving
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiao, Chang, Yang, Brenda |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LLMs May Not Be Human-Level Players, But They Can Be Testers: Measuring Game Difficulty with LLM Agents
por: Xiao, Chang, et al.
Publicado: (2024)
por: Xiao, Chang, et al.
Publicado: (2024)
An Efficient and Streaming Audio Visual Active Speaker Detection System
por: Kundu, Arnav, et al.
Publicado: (2024)
por: Kundu, Arnav, et al.
Publicado: (2024)
Exploring Eye Tracking to Detect Cognitive Load in Complex Virtual Reality Training
por: Nasri, Mahsa, et al.
Publicado: (2024)
por: Nasri, Mahsa, et al.
Publicado: (2024)
Quantifying Visual Properties of GAM Shape Plots: Impact on Perceived Cognitive Load and Interpretability
por: Kruschel, Sven, et al.
Publicado: (2024)
por: Kruschel, Sven, et al.
Publicado: (2024)
Multi-Domain EEG Representation Learning with Orthogonal Mapping and Attention-based Fusion for Cognitive Load Classification
por: Angkan, Prithila, et al.
Publicado: (2025)
por: Angkan, Prithila, et al.
Publicado: (2025)
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion
por: Vu, Evgeniia, et al.
Publicado: (2025)
por: Vu, Evgeniia, et al.
Publicado: (2025)
Neuro-Informed Joint Learning Enhances Cognitive Workload Decoding in Portable BCIs
por: Yang, Xiaoxiao, et al.
Publicado: (2025)
por: Yang, Xiaoxiao, et al.
Publicado: (2025)
Mazed and Confused: A Dataset of Cybersickness, Working Memory, Mental Load, Physical Load, and Attention During a Real Walking Task in VR
por: Setu, Jyotirmay Nag, et al.
Publicado: (2024)
por: Setu, Jyotirmay Nag, et al.
Publicado: (2024)
Semi-Supervised Dual-Stream Self-Attentive Adversarial Graph Contrastive Learning for Cross-Subject EEG-based Emotion Recognition
por: Ye, Weishan, et al.
Publicado: (2023)
por: Ye, Weishan, et al.
Publicado: (2023)
Chronotome: Real-Time Topic Modeling for Streaming Embedding Spaces
por: Lim, Matte, et al.
Publicado: (2025)
por: Lim, Matte, et al.
Publicado: (2025)
Mitigating Cognitive Biases in Multi-Criteria Crowd Assessment
por: Ito, Shun, et al.
Publicado: (2024)
por: Ito, Shun, et al.
Publicado: (2024)
Do Recommender Systems Promote Local Music? A Reproducibility Study Using Music Streaming Data
por: Matrosova, Kristina, et al.
Publicado: (2024)
por: Matrosova, Kristina, et al.
Publicado: (2024)
LLM-based Cognitive Models of Students with Misconceptions
por: Sonkar, Shashank, et al.
Publicado: (2024)
por: Sonkar, Shashank, et al.
Publicado: (2024)
Will You Be Aware? Eye Tracking-Based Modeling of Situational Awareness in Augmented Reality
por: Qu, Zhehan, et al.
Publicado: (2025)
por: Qu, Zhehan, et al.
Publicado: (2025)
ECVL-ROUTER: Scenario-Aware Routing for Vision-Language Models
por: Tang, Xin, et al.
Publicado: (2025)
por: Tang, Xin, et al.
Publicado: (2025)
Cognitive State Inference from VR Motion via Motion Foundation Model
por: Wen, Kaiang, et al.
Publicado: (2025)
por: Wen, Kaiang, et al.
Publicado: (2025)
FUnc-SNE: A flexible, Fast, and Unconstrained algorithm for neighbour embeddings
por: Lambert, Pierre, et al.
Publicado: (2025)
por: Lambert, Pierre, et al.
Publicado: (2025)
Hybrid Deep Learning Model to Estimate Cognitive Effort from fNIRS Signals
por: Sharmin, Shayla, et al.
Publicado: (2025)
por: Sharmin, Shayla, et al.
Publicado: (2025)
Cognitive Energy Modeling for Neuroadaptive Human-Machine Systems using EEG and WGAN-GP
por: Sattiraju, Sriram, et al.
Publicado: (2026)
por: Sattiraju, Sriram, et al.
Publicado: (2026)
LLM Chatbot-Creation Approaches
por: Mehta, Hemil, et al.
Publicado: (2025)
por: Mehta, Hemil, et al.
Publicado: (2025)
VoXtream: Full-Stream Text-to-Speech with Extremely Low Latency
por: Torgashov, Nikita, et al.
Publicado: (2025)
por: Torgashov, Nikita, et al.
Publicado: (2025)
Cost-Aware Bayesian Optimization for Prototyping Interactive Devices
por: Langerak, Thomas, et al.
Publicado: (2026)
por: Langerak, Thomas, et al.
Publicado: (2026)
Hardware-Aware Federated Learning for Speech Emotion Recognition
por: Yuksel, Beyazit Bestami, et al.
Publicado: (2026)
por: Yuksel, Beyazit Bestami, et al.
Publicado: (2026)
ML Mule: Mobile-Driven Context-Aware Collaborative Learning
por: Yu, Haoxiang, et al.
Publicado: (2025)
por: Yu, Haoxiang, et al.
Publicado: (2025)
Be There, Be Together, Be Streamed! AR Scenic Live-Streaming for an Interactive and Collective Experience
por: Huang, Zeyu, et al.
Publicado: (2024)
por: Huang, Zeyu, et al.
Publicado: (2024)
Distortion-Aware Brushing for Reliable Cluster Analysis in Multidimensional Projections
por: Jeon, Hyeon, et al.
Publicado: (2022)
por: Jeon, Hyeon, et al.
Publicado: (2022)
UR2M: Uncertainty and Resource-Aware Event Detection on Microcontrollers
por: Jia, Hong, et al.
Publicado: (2024)
por: Jia, Hong, et al.
Publicado: (2024)
Can LLM Assist in the Evaluation of the Quality of Machine Learning Explanations?
por: Wang, Bo, et al.
Publicado: (2025)
por: Wang, Bo, et al.
Publicado: (2025)
Evaluation of LLM-based Explanations for a Learning Analytics Dashboard
por: Deriyeva, Alina, et al.
Publicado: (2025)
por: Deriyeva, Alina, et al.
Publicado: (2025)
Geometry-Aware Active Learning of Pattern Rankings via Choquet-Based Aggregation
por: Opran, Tudor Matei, et al.
Publicado: (2025)
por: Opran, Tudor Matei, et al.
Publicado: (2025)
HappyRouting: Learning Emotion-Aware Route Trajectories for Scalable In-The-Wild Navigation
por: Bethge, David, et al.
Publicado: (2024)
por: Bethge, David, et al.
Publicado: (2024)
Efficient Online Crowdsourcing with Complex Annotations
por: Meir, Reshef, et al.
Publicado: (2024)
por: Meir, Reshef, et al.
Publicado: (2024)
CataractBot: An LLM-Powered Expert-in-the-Loop Chatbot for Cataract Patients
por: Ramjee, Pragnya, et al.
Publicado: (2024)
por: Ramjee, Pragnya, et al.
Publicado: (2024)
Data-Prompt Co-Evolution: Growing Test Sets to Refine LLM Behavior
por: Lee, Minjae, et al.
Publicado: (2025)
por: Lee, Minjae, et al.
Publicado: (2025)
Uncertainty-Aware Cross-Modal Knowledge Distillation with Prototype Learning for Multimodal Brain-Computer Interfaces
por: Jang, Hyo-Jeong, et al.
Publicado: (2025)
por: Jang, Hyo-Jeong, et al.
Publicado: (2025)
Emotion-Aware Interaction Design in Intelligent User Interface Using Multi-Modal Deep Learning
por: Duan, Shiyu, et al.
Publicado: (2024)
por: Duan, Shiyu, et al.
Publicado: (2024)
CTG-Insight: A Multi-Agent Interpretable LLM Framework for Cardiotocography Analysis and Classification
por: Sun, Black, et al.
Publicado: (2025)
por: Sun, Black, et al.
Publicado: (2025)
Private Yet Social: How LLM Chatbots Support and Challenge Eating Disorder Recovery
por: Choi, Ryuhaerang, et al.
Publicado: (2024)
por: Choi, Ryuhaerang, et al.
Publicado: (2024)
LangLasso: Interactive Cluster Descriptions through LLM Explanation
por: Buchmüller, Raphael, et al.
Publicado: (2026)
por: Buchmüller, Raphael, et al.
Publicado: (2026)
The Role of Visualization in LLM-Assisted Knowledge Graph Systems: Effects on User Trust, Exploration, and Workflows
por: Li, Harry, et al.
Publicado: (2025)
por: Li, Harry, et al.
Publicado: (2025)
Ejemplares similares
-
LLMs May Not Be Human-Level Players, But They Can Be Testers: Measuring Game Difficulty with LLM Agents
por: Xiao, Chang, et al.
Publicado: (2024) -
An Efficient and Streaming Audio Visual Active Speaker Detection System
por: Kundu, Arnav, et al.
Publicado: (2024) -
Exploring Eye Tracking to Detect Cognitive Load in Complex Virtual Reality Training
por: Nasri, Mahsa, et al.
Publicado: (2024) -
Quantifying Visual Properties of GAM Shape Plots: Impact on Perceived Cognitive Load and Interpretability
por: Kruschel, Sven, et al.
Publicado: (2024) -
Multi-Domain EEG Representation Learning with Orthogonal Mapping and Attention-based Fusion for Cognitive Load Classification
por: Angkan, Prithila, et al.
Publicado: (2025)