HealthSLM-Bench: Benchmarking Small Language Models for Mobile and Wearable Healthcare Monitoring
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Xin, Dang, Ting, Zhang, Xinyu, Kostakos, Vassilis, Witbrock, Michael J., Jia, Hong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Efficient and Personalized Mobile Health Event Prediction via Small Language Models
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
ScreenTK: Seamless Detection of Time-Killing Moments Using Continuous Mobile Screen Text and On-Device LLMs
by: Fang, Le, et al.
Published: (2024)
by: Fang, Le, et al.
Published: (2024)
Do LLMs Need to See Everything? A Benchmark and Study of Failures in LLM-driven Smartphone Automation using Screentext vs. Screenshots
by: Zhang, Shiquan, et al.
Published: (2026)
by: Zhang, Shiquan, et al.
Published: (2026)
Predicting Affective States from Screen Text Sentiment
by: Teng, Songyan, et al.
Published: (2024)
by: Teng, Songyan, et al.
Published: (2024)
Enabling On-Device LLMs Personalization with Smartphone Sensing
by: Zhang, Shiquan, et al.
Published: (2024)
by: Zhang, Shiquan, et al.
Published: (2024)
Situationally-Induced Impairments and Disabilities Research
by: Sarsenbayeva, Zhanna, et al.
Published: (2019)
by: Sarsenbayeva, Zhanna, et al.
Published: (2019)
MobileAgentBench: An Efficient and User-Friendly Benchmark for Mobile LLM Agents
by: Wang, Luyuan, et al.
Published: (2024)
by: Wang, Luyuan, et al.
Published: (2024)
Real-Time Detection of Robot Failures Using Gaze Dynamics in Collaborative Tasks
by: Tabatabaei, Ramtin, et al.
Published: (2025)
by: Tabatabaei, Ramtin, et al.
Published: (2025)
Gazing at Failure: Investigating Human Gaze in Response to Robot Failure in Collaborative Tasks
by: Tabatabaei, Ramtin, et al.
Published: (2025)
by: Tabatabaei, Ramtin, et al.
Published: (2025)
AutoJournaling: A Context-Aware Journaling System Leveraging MLLMs on Smartphone Screenshots
by: Zhang, Tianyi, et al.
Published: (2024)
by: Zhang, Tianyi, et al.
Published: (2024)
VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics
by: Gong, Yichen, et al.
Published: (2026)
by: Gong, Yichen, et al.
Published: (2026)
AmbiBench: Benchmarking Mobile GUI Agents Beyond One-Shot Instructions in the Wild
by: Sun, Jiazheng, et al.
Published: (2026)
by: Sun, Jiazheng, et al.
Published: (2026)
Wearable Healthcare Devices for Monitoring Stress and Attention Level in Workplace Environments
by: Traunmuller, Peter, et al.
Published: (2024)
by: Traunmuller, Peter, et al.
Published: (2024)
Exploring Large-Scale Language Models to Evaluate EEG-Based Multimodal Data for Mental Health
by: Hu, Yongquan, et al.
Published: (2024)
by: Hu, Yongquan, et al.
Published: (2024)
VitalAgent: A Tool-Augmented Agent for Reactive and Proactive Physiological Monitoring over Wearable Health Data
by: Zhu, Di, et al.
Published: (2026)
by: Zhu, Di, et al.
Published: (2026)
From Patient Burdens to User Agency: Designing for Real-Time Protection Support in Online Health Consultations
by: Zhang, Shuning, et al.
Published: (2025)
by: Zhang, Shuning, et al.
Published: (2025)
WearBCI Dataset: Understanding and Benchmarking Real-World Wearable Brain-Computer Interfaces Signals
by: Liu, Haoxian, et al.
Published: (2026)
by: Liu, Haoxian, et al.
Published: (2026)
ProMemAssist: Exploring Timely Proactive Assistance Through Working Memory Modeling in Multi-Modal Wearable Devices
by: Pu, Kevin, et al.
Published: (2025)
by: Pu, Kevin, et al.
Published: (2025)
Heart2Mind: Human-Centered Contestable Psychiatric Disorder Diagnosis System using Wearable ECG Monitors
by: Nguyen, Hung, et al.
Published: (2025)
by: Nguyen, Hung, et al.
Published: (2025)
DisaBench: A Participatory Evaluation Framework for Disability Harms in Language Models
by: Kim, Eugenia, et al.
Published: (2026)
by: Kim, Eugenia, et al.
Published: (2026)
Large Language Models for Wearable Sensor-Based Human Activity Recognition, Health Monitoring, and Behavioral Modeling: A Survey of Early Trends, Datasets, and Challenges
by: Ferrara, Emilio
Published: (2024)
by: Ferrara, Emilio
Published: (2024)
An Empirical Evaluation of AI-Powered Non-Player Characters' Perceived Realism and Performance in Virtual Reality Environments
by: Korkiakoski, Mikko, et al.
Published: (2025)
by: Korkiakoski, Mikko, et al.
Published: (2025)
Scaling Wearable Foundation Models
by: Narayanswamy, Girish, et al.
Published: (2024)
by: Narayanswamy, Girish, et al.
Published: (2024)
PhysioLLM: Supporting Personalized Health Insights with Wearables and Large Language Models
by: Fang, Cathy Mengying, et al.
Published: (2024)
by: Fang, Cathy Mengying, et al.
Published: (2024)
AquaVLM: Improving Underwater Situation Awareness with Mobile Vision Language Models
by: Tian, Beitong, et al.
Published: (2025)
by: Tian, Beitong, et al.
Published: (2025)
InfoPrint: Embedding Information into 3D Printed Objects
by: Jiang, Weiwei, et al.
Published: (2021)
by: Jiang, Weiwei, et al.
Published: (2021)
SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents
by: Ai, Kuangshi, et al.
Published: (2026)
by: Ai, Kuangshi, et al.
Published: (2026)
FingerTip 20K: A Benchmark for Proactive and Personalized Mobile LLM Agents
by: Yang, Qinglong, et al.
Published: (2025)
by: Yang, Qinglong, et al.
Published: (2025)
EEG-FM-Bench: A Comprehensive Benchmark for the Systematic Evaluation of EEG Foundation Models
by: Xiong, Wei, et al.
Published: (2025)
by: Xiong, Wei, et al.
Published: (2025)
MindBenchAI: An Actionable Platform to Evaluate the Profile and Performance of Large Language Models in a Mental Healthcare Context
by: Dwyer, Bridget, et al.
Published: (2025)
by: Dwyer, Bridget, et al.
Published: (2025)
Wearable Meets LLM for Stress Management: A Duoethnographic Study Integrating Wearable-Triggered Stressors and LLM Chatbots for Personalized Interventions
by: Neupane, Sameer, et al.
Published: (2025)
by: Neupane, Sameer, et al.
Published: (2025)
A Foundation Model for Wearable Movement Data in Mental Health Research
by: Ruan, Franklin Y., et al.
Published: (2024)
by: Ruan, Franklin Y., et al.
Published: (2024)
ScreenAudit: Detecting Screen Reader Accessibility Errors in Mobile Apps Using Large Language Models
by: Zhong, Mingyuan, et al.
Published: (2025)
by: Zhong, Mingyuan, et al.
Published: (2025)
MHDash: An Online Platform for Benchmarking Mental Health-Aware AI Assistants
by: Zhang, Yihe, et al.
Published: (2026)
by: Zhang, Yihe, et al.
Published: (2026)
TOM: A Development Platform For Wearable Intelligent Assistants
by: Janaka, Nuwan, et al.
Published: (2024)
by: Janaka, Nuwan, et al.
Published: (2024)
Beyond Permissions: Investigating Mobile Personalization with Simulated Personas
by: Khalilov, Ibrahim, et al.
Published: (2025)
by: Khalilov, Ibrahim, et al.
Published: (2025)
A Scoping Review of AI-Driven Digital Interventions in Mental Health Care: Mapping Applications Across Screening, Support, Monitoring, Prevention, and Clinical Education
by: Ni, Yang, et al.
Published: (2026)
by: Ni, Yang, et al.
Published: (2026)
Wearable Device-Based Real-Time Monitoring of Physiological Signals: Evaluating Cognitive Load Across Different Tasks
by: He, Ling, et al.
Published: (2024)
by: He, Ling, et al.
Published: (2024)
Large Language Models in Peer-Run Community Behavioral Health Services: Understanding Peer Specialists and Service Users' Perspectives on Opportunities, Risks, and Mitigation Strategies
by: Peng, Cindy, et al.
Published: (2026)
by: Peng, Cindy, et al.
Published: (2026)
A Scalable Framework for Evaluating Health Language Models
by: Mallinar, Neil, et al.
Published: (2025)
by: Mallinar, Neil, et al.
Published: (2025)
Similar Items
-
Efficient and Personalized Mobile Health Event Prediction via Small Language Models
by: Wang, Xin, et al.
Published: (2024) -
ScreenTK: Seamless Detection of Time-Killing Moments Using Continuous Mobile Screen Text and On-Device LLMs
by: Fang, Le, et al.
Published: (2024) -
Do LLMs Need to See Everything? A Benchmark and Study of Failures in LLM-driven Smartphone Automation using Screentext vs. Screenshots
by: Zhang, Shiquan, et al.
Published: (2026) -
Predicting Affective States from Screen Text Sentiment
by: Teng, Songyan, et al.
Published: (2024) -
Enabling On-Device LLMs Personalization with Smartphone Sensing
by: Zhang, Shiquan, et al.
Published: (2024)