Identification of Cognitive Decline from Spoken Language through Feature Selection and the Bag of Acoustic Words Model
Fuente:
arXiv
Guardado en:
| Autores principales: | Niemelä, Marko, von Bonsdorff, Mikaela, Äyrämö, Sami, Kärkkäinen, Tommi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Dementia classification from spontaneous speech using wrapper-based feature selection
por: Niemelä, Marko, et al.
Publicado: (2025)
por: Niemelä, Marko, et al.
Publicado: (2025)
On the use of Performer and Agent Attention for Spoken Language Identification
por: dhiman, Jitendra Kumar, et al.
Publicado: (2025)
por: dhiman, Jitendra Kumar, et al.
Publicado: (2025)
Abusive Speech Detection in Indic Languages Using Acoustic Features
por: Spiesberger, Anika A., et al.
Publicado: (2024)
por: Spiesberger, Anika A., et al.
Publicado: (2024)
E-chat: Emotion-sensitive Spoken Dialogue System with Large Language Models
por: Xue, Hongfei, et al.
Publicado: (2023)
por: Xue, Hongfei, et al.
Publicado: (2023)
BERT-LID: Leveraging BERT to Improve Spoken Language Identification
por: Nie, Yuting, et al.
Publicado: (2022)
por: Nie, Yuting, et al.
Publicado: (2022)
IMPACT: Industrial Machine Perception via Acoustic Cognitive Transformer
por: Han, Changheon, et al.
Publicado: (2025)
por: Han, Changheon, et al.
Publicado: (2025)
CoPlay: Audio-agnostic Cognitive Scaling for Acoustic Sensing
por: Li, Yin, et al.
Publicado: (2024)
por: Li, Yin, et al.
Publicado: (2024)
Keyword Mamba: Spoken Keyword Spotting with State Space Models
por: Ding, Hanyu, et al.
Publicado: (2025)
por: Ding, Hanyu, et al.
Publicado: (2025)
SAMOS: A Neural MOS Prediction Model Leveraging Semantic Representations and Acoustic Features
por: Shi, Yu-Fei, et al.
Publicado: (2024)
por: Shi, Yu-Fei, et al.
Publicado: (2024)
WHISMA: A Speech-LLM to Perform Zero-shot Spoken Language Understanding
por: Li, Mohan, et al.
Publicado: (2024)
por: Li, Mohan, et al.
Publicado: (2024)
LALM-as-a-Judge: Benchmarking Large Audio-Language Models for Safety Evaluation in Multi-Turn Spoken Dialogues
por: Ivry, Amir, et al.
Publicado: (2026)
por: Ivry, Amir, et al.
Publicado: (2026)
Identification of Physical Properties in Acoustic Tubes Using Physics-Informed Neural Networks
por: Yokota, Kazuya, et al.
Publicado: (2024)
por: Yokota, Kazuya, et al.
Publicado: (2024)
Comparative Evaluation of Acoustic Feature Extraction Tools for Clinical Speech Analysis
por: Choi, Anna Seo Gyeong, et al.
Publicado: (2025)
por: Choi, Anna Seo Gyeong, et al.
Publicado: (2025)
Semantic-Aware Interruption Detection in Spoken Dialogue Systems: Benchmark, Metric, and Model
por: Xia, Kangxiang, et al.
Publicado: (2026)
por: Xia, Kangxiang, et al.
Publicado: (2026)
Exploring Spoken Language Identification Strategies for Automatic Transcription of Multilingual Broadcast and Institutional Speech
por: Valente, Martina, et al.
Publicado: (2024)
por: Valente, Martina, et al.
Publicado: (2024)
J-CHAT: Japanese Large-scale Spoken Dialogue Corpus for Spoken Dialogue Language Modeling
por: Nakata, Wataru, et al.
Publicado: (2024)
por: Nakata, Wataru, et al.
Publicado: (2024)
SD-Eval: A Benchmark Dataset for Spoken Dialogue Understanding Beyond Words
por: Ao, Junyi, et al.
Publicado: (2024)
por: Ao, Junyi, et al.
Publicado: (2024)
VAE-based Phoneme Alignment Using Gradient Annealing and SSL Acoustic Features
por: Koriyama, Tomoki
Publicado: (2024)
por: Koriyama, Tomoki
Publicado: (2024)
Improving Acoustic Word Embeddings through Correspondence Training of Self-supervised Speech Representations
por: Meghanani, Amit, et al.
Publicado: (2024)
por: Meghanani, Amit, et al.
Publicado: (2024)
VQTTS: High-Fidelity Text-to-Speech Synthesis with Self-Supervised VQ Acoustic Feature
por: Du, Chenpeng, et al.
Publicado: (2022)
por: Du, Chenpeng, et al.
Publicado: (2022)
Spoken language change detection inspired by speaker change detection
por: Mishra, Jagabandhu, et al.
Publicado: (2023)
por: Mishra, Jagabandhu, et al.
Publicado: (2023)
On The Landscape of Spoken Language Models: A Comprehensive Survey
por: Arora, Siddhant, et al.
Publicado: (2025)
por: Arora, Siddhant, et al.
Publicado: (2025)
Spirit LM: Interleaved Spoken and Written Language Model
por: Nguyen, Tu Anh, et al.
Publicado: (2024)
por: Nguyen, Tu Anh, et al.
Publicado: (2024)
On the Evaluation of Speech Foundation Models for Spoken Language Understanding
por: Arora, Siddhant, et al.
Publicado: (2024)
por: Arora, Siddhant, et al.
Publicado: (2024)
Long-Form Speech Generation with Spoken Language Models
por: Park, Se Jin, et al.
Publicado: (2024)
por: Park, Se Jin, et al.
Publicado: (2024)
Integrating Pause Information with Word Embeddings in Language Models for Alzheimer's Disease Detection from Spontaneous Speech
por: Pu, Yu, et al.
Publicado: (2025)
por: Pu, Yu, et al.
Publicado: (2025)
Listen through the Sound: Generative Speech Restoration Leveraging Acoustic Context Representation
por: Chung, Soo-Whan, et al.
Publicado: (2025)
por: Chung, Soo-Whan, et al.
Publicado: (2025)
Frequency Tracking Features for Data-Efficient Deep Siren Identification
por: Damiano, Stefano, et al.
Publicado: (2024)
por: Damiano, Stefano, et al.
Publicado: (2024)
Improving Language Model-Based Zero-Shot Text-to-Speech Synthesis with Multi-Scale Acoustic Prompts
por: Lei, Shun, et al.
Publicado: (2023)
por: Lei, Shun, et al.
Publicado: (2023)
Kanade: A Simple Disentangled Tokenizer for Spoken Language Modeling
por: Huang, Zhijie, et al.
Publicado: (2026)
por: Huang, Zhijie, et al.
Publicado: (2026)
Are These Even Words? Quantifying the Gibberishness of Generative Speech Models
por: de Oliveira, Danilo, et al.
Publicado: (2025)
por: de Oliveira, Danilo, et al.
Publicado: (2025)
Evaluating Hallucinations in Audio-Visual Multimodal LLMs with Spoken Queries under Diverse Acoustic Conditions
por: Park, Hansol, et al.
Publicado: (2025)
por: Park, Hansol, et al.
Publicado: (2025)
Frequency & Channel Attention Network for Small Footprint Noisy Spoken Keyword Spotting
por: Lin, Yuanxi, et al.
Publicado: (2024)
por: Lin, Yuanxi, et al.
Publicado: (2024)
Evaluation of Virtual Acoustic Environments with Different Acoustic Level of Detail
por: Fichna, Stefan, et al.
Publicado: (2023)
por: Fichna, Stefan, et al.
Publicado: (2023)
Streaming Endpointer for Spoken Dialogue using Neural Audio Codecs and Label-Delayed Training
por: Udupa, Sathvik, et al.
Publicado: (2025)
por: Udupa, Sathvik, et al.
Publicado: (2025)
TASTE: Text-Aligned Speech Tokenization and Embedding for Spoken Language Modeling
por: Tseng, Liang-Hsuan, et al.
Publicado: (2025)
por: Tseng, Liang-Hsuan, et al.
Publicado: (2025)
Acoustical Features as Knee Health Biomarkers: A Critical Analysis
por: Kechris, Christodoulos, et al.
Publicado: (2024)
por: Kechris, Christodoulos, et al.
Publicado: (2024)
Voice Conversion Improves Cross-Domain Robustness for Spoken Arabic Dialect Identification
por: Abdullah, Badr M., et al.
Publicado: (2025)
por: Abdullah, Badr M., et al.
Publicado: (2025)
Unraveling Complex Data Diversity in Underwater Acoustic Target Recognition through Convolution-based Mixture of Experts
por: Xie, Yuan, et al.
Publicado: (2024)
por: Xie, Yuan, et al.
Publicado: (2024)
Implicit Self-supervised Language Representation for Spoken Language Diarization
por: Mishra, Jagabandhu, et al.
Publicado: (2023)
por: Mishra, Jagabandhu, et al.
Publicado: (2023)
Ejemplares similares
-
Dementia classification from spontaneous speech using wrapper-based feature selection
por: Niemelä, Marko, et al.
Publicado: (2025) -
On the use of Performer and Agent Attention for Spoken Language Identification
por: dhiman, Jitendra Kumar, et al.
Publicado: (2025) -
Abusive Speech Detection in Indic Languages Using Acoustic Features
por: Spiesberger, Anika A., et al.
Publicado: (2024) -
E-chat: Emotion-sensitive Spoken Dialogue System with Large Language Models
por: Xue, Hongfei, et al.
Publicado: (2023) -
BERT-LID: Leveraging BERT to Improve Spoken Language Identification
por: Nie, Yuting, et al.
Publicado: (2022)