Using LLMs for Late Multimodal Sensor Fusion for Activity Recognition
Fuente:
arXiv
Saved in:
| Main Authors: | Demirel, Ilker, Thakkar, Karan, Elizalde, Benjamin, Marques, Miquel Espi, Sarathy, Aditya, Bai, Yang, Srinivas, Umamahesh, Xu, Jiajie, Ren, Shirley, Narain, Jaya |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Speech Foundation Models Generalize to Time Series Tasks from Wearable Sensor Data
by: Narain, Jaya, et al.
Published: (2025)
by: Narain, Jaya, et al.
Published: (2025)
Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect
by: Narain, Jaya, et al.
Published: (2025)
by: Narain, Jaya, et al.
Published: (2025)
Affect Models Have Weak Generalizability to Atypical Speech
by: Narain, Jaya, et al.
Published: (2025)
by: Narain, Jaya, et al.
Published: (2025)
DECAF: Dynamic Envelope Context-Aware Fusion for Speech-Envelope Reconstruction from EEG
by: Thakkar, Karan, et al.
Published: (2026)
by: Thakkar, Karan, et al.
Published: (2026)
Unimodal and Multimodal Sensor Fusion for Wearable Activity Recognition
by: Bello, Hymalai
Published: (2024)
by: Bello, Hymalai
Published: (2024)
Do LLMs estimate uncertainty well in instruction-following?
by: Heo, Juyeon, et al.
Published: (2024)
by: Heo, Juyeon, et al.
Published: (2024)
Do LLMs "know" internally when they follow instructions?
by: Heo, Juyeon, et al.
Published: (2024)
by: Heo, Juyeon, et al.
Published: (2024)
LLMs can construct powerful representations and streamline sample-efficient supervised learning
by: Demirel, Ilker, et al.
Published: (2026)
by: Demirel, Ilker, et al.
Published: (2026)
Multimodal Fusion and Interpretability in Human Activity Recognition: A Reproducible Framework for Sensor-Based Modeling
by: Yang, Yiyao, et al.
Published: (2025)
by: Yang, Yiyao, et al.
Published: (2025)
ECONOMIC DEVELOPMENT AND MARKETING STRATEGIES: A COMPARATIVE LENS
by: Ravi Sarathy
Published: (2014)
by: Ravi Sarathy
Published: (2014)
Triple Spectral Fusion for Sensor-based Human Activity Recognition
by: Zhang, Ye, et al.
Published: (2026)
by: Zhang, Ye, et al.
Published: (2026)
Chain-of-Thought Degrades Visual Spatial Reasoning Capabilities of Multimodal LLMs
by: Kancheti, Sai Srinivas, et al.
Published: (2026)
by: Kancheti, Sai Srinivas, et al.
Published: (2026)
¿Tiene solución la ciudad?
by: Mariano Vázquez Espí
Published: (2008)
by: Mariano Vázquez Espí
Published: (2008)
PERCEPCIÓN AFECTIVA DEL ALUMNADO EN EDUCACIÓN FÍSICA PRIMARIA EN LA PROVINCIA DE ALICANTE
by: Andrea Jordá-Espi
Published: (2019)
by: Andrea Jordá-Espi
Published: (2019)
Construcciones utópicas: tres tesis y una regla práctica
by: Mariano Vázquez Espí
Published: (2003)
by: Mariano Vázquez Espí
Published: (2003)
Prediction-powered Generalization of Causal Inferences
by: Demirel, Ilker, et al.
Published: (2024)
by: Demirel, Ilker, et al.
Published: (2024)
Federated Multi-Armed Bandits Under Byzantine Attacks
by: Saday, Artun, et al.
Published: (2022)
by: Saday, Artun, et al.
Published: (2022)
Virtual Fusion with Contrastive Learning for Single Sensor-based Activity Recognition
by: Nguyen, Duc-Anh, et al.
Published: (2023)
by: Nguyen, Duc-Anh, et al.
Published: (2023)
SpatialGeo:Boosting Spatial Reasoning in Multimodal LLMs via Geometry-Semantics Fusion
by: Guo, Jiajie, et al.
Published: (2025)
by: Guo, Jiajie, et al.
Published: (2025)
Efficient Vocabulary-Free Fine-Grained Visual Recognition in the Age of Multimodal LLMs
by: Kuchibhotla, Hari Chandana, et al.
Published: (2025)
by: Kuchibhotla, Hari Chandana, et al.
Published: (2025)
Fusion and Grouping Strategies in Deep Learning for Local Climate Zone Classification of Multimodal Remote Sensing Data
by: Thomas, Ancymol, et al.
Published: (2026)
by: Thomas, Ancymol, et al.
Published: (2026)
Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs
by: Sinha, Rohit, et al.
Published: (2026)
by: Sinha, Rohit, et al.
Published: (2026)
Design and Development of Laughter Recognition System Based on Multimodal Fusion and Deep Learning
by: Zhao, Fuzheng, et al.
Published: (2024)
by: Zhao, Fuzheng, et al.
Published: (2024)
Uncovering Bias Mechanisms in Observational Studies
by: Demirel, Ilker, et al.
Published: (2025)
by: Demirel, Ilker, et al.
Published: (2025)
ANÁLISIS DE LAS FALACIAS EN TORNO A LA TEORÍA DE LA SOBERANÍA NACIONAL (O POPULAR)
by: Luis-Tomás Zapater Espí
Published: (2017)
by: Luis-Tomás Zapater Espí
Published: (2017)
La visión del estudiante respecto a la Diplomatura de Fisioterapia en la Universitat de València: un estudio descriptivo
by: Gemma Victoria Espí López
Published: (2012)
by: Gemma Victoria Espí López
Published: (2012)
Cross-Level Sensor Fusion with Object Lists via Transformer for 3D Object Detection
by: Liu, Xiangzhong, et al.
Published: (2025)
by: Liu, Xiangzhong, et al.
Published: (2025)
Post Fusion Bird's Eye View Feature Stabilization for Robust Multimodal 3D Detection
by: Dong, Trung Tien, et al.
Published: (2026)
by: Dong, Trung Tien, et al.
Published: (2026)
Tools in the Loop: Quantifying Uncertainty of LLM Question Answering Systems That Use Tools
by: Lymperopoulos, Panagiotis, et al.
Published: (2025)
by: Lymperopoulos, Panagiotis, et al.
Published: (2025)
Differential Perspectives: Epistemic Disconnects Surrounding the US Census Bureau's Use of Differential Privacy
by: boyd, danah, et al.
Published: (2026)
by: boyd, danah, et al.
Published: (2026)
Centering Policy and Practice: Research Gaps around Usable Differential Privacy
by: Cummings, Rachel, et al.
Published: (2024)
by: Cummings, Rachel, et al.
Published: (2024)
Statistical Imaginaries, State Legitimacy: Grappling with the Arrangements Underpinning Quantification in the U.S. Census
by: Sarathy, Jayshree, et al.
Published: (2026)
by: Sarathy, Jayshree, et al.
Published: (2026)
Analyzing the Differentially Private Theil-Sen Estimator for Simple Linear Regression
by: Sarathy, Jayshree, et al.
Published: (2022)
by: Sarathy, Jayshree, et al.
Published: (2022)
Analogical Reasoning Within a Conceptual Hyperspace
by: Goldowsky, Howard, et al.
Published: (2024)
by: Goldowsky, Howard, et al.
Published: (2024)
Finding Needles in Images: Can Multimodal LLMs Locate Fine Details?
by: Thakkar, Parth, et al.
Published: (2025)
by: Thakkar, Parth, et al.
Published: (2025)
RelCon: Relative Contrastive Learning for a Motion Foundation Model for Wearable Data
by: Xu, Maxwell A., et al.
Published: (2024)
by: Xu, Maxwell A., et al.
Published: (2024)
The Representation of a Logo in Memory: A Study of Recall and Recognition of the Google Logo
by: Medine Elif Demirel, et al.
Published: (2025)
by: Medine Elif Demirel, et al.
Published: (2025)
El federalismo de la India está plagado de altercados por las vías fluviales
by: Narain, A
Published: (2009)
by: Narain, A
Published: (2009)
Water Security, Conflict and Cooperation in Peri-Urban South Asia Flows across Boundaries
by: Vishal Narain
by: Vishal Narain
Statistical genomics and bioinformatics
by: Prem Narain
Published: (2010)
by: Prem Narain
Published: (2010)
Similar Items
-
Speech Foundation Models Generalize to Time Series Tasks from Wearable Sensor Data
by: Narain, Jaya, et al.
Published: (2025) -
Voice Quality Dimensions as Interpretable Primitives for Speaking Style for Atypical Speech and Affect
by: Narain, Jaya, et al.
Published: (2025) -
Affect Models Have Weak Generalizability to Atypical Speech
by: Narain, Jaya, et al.
Published: (2025) -
DECAF: Dynamic Envelope Context-Aware Fusion for Speech-Envelope Reconstruction from EEG
by: Thakkar, Karan, et al.
Published: (2026) -
Unimodal and Multimodal Sensor Fusion for Wearable Activity Recognition
by: Bello, Hymalai
Published: (2024)