A Dual-Stage Time-Context Network for Speech-Based Alzheimer's Disease Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Yifan, Guo, Long, Liu, Hong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Leveraging Multimodal Methods and Spontaneous Speech for Alzheimer's Disease Identification
von: Gao, Yifan, et al.
Veröffentlicht: (2024)
von: Gao, Yifan, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models for Spontaneous Speech-Based Suicide Risk Detection
von: Gao, Yifan, et al.
Veröffentlicht: (2025)
von: Gao, Yifan, et al.
Veröffentlicht: (2025)
Leveraging Prompt Learning and Pause Encoding for Alzheimer's Disease Detection
von: Liu, Yin-Long, et al.
Veröffentlicht: (2024)
von: Liu, Yin-Long, et al.
Veröffentlicht: (2024)
Exploring Gender Bias in Alzheimer's Disease Detection: Insights from Mandarin and Greek Speech Perception
von: He, Liu, et al.
Veröffentlicht: (2025)
von: He, Liu, et al.
Veröffentlicht: (2025)
Two-Stage Acoustic Adaptation with Gated Cross-Attention Adapters for LLM-Based Multi-Talker Speech Recognition
von: Shi, Hao, et al.
Veröffentlicht: (2026)
von: Shi, Hao, et al.
Veröffentlicht: (2026)
Spectral Masking with Explicit Time-Context Windowing for Neural Network-Based Monaural Speech Enhancement
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2024)
von: Fiorio, Luan Vinícius, et al.
Veröffentlicht: (2024)
DS-Codec: Dual-Stage Training with Mirror-to-NonMirror Architecture Switching for Speech Codec
von: Chen, Peijie, et al.
Veröffentlicht: (2025)
von: Chen, Peijie, et al.
Veröffentlicht: (2025)
Conditional Latent Diffusion-Based Speech Enhancement Via Dual Context Learning
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
von: Zhao, Shengkui, et al.
Veröffentlicht: (2025)
Online Register for Dual-Mode Self-Supervised Speech Models: Mitigating The Lack of Future Context
von: Goto, Keita, et al.
Veröffentlicht: (2026)
von: Goto, Keita, et al.
Veröffentlicht: (2026)
A Multi-Stage Framework for Multimodal Controllable Speech Synthesis
von: Niu, Rui, et al.
Veröffentlicht: (2025)
von: Niu, Rui, et al.
Veröffentlicht: (2025)
RTCFake: Speech Deepfake Detection in Real-Time Communication
von: Xue, Jun, et al.
Veröffentlicht: (2026)
von: Xue, Jun, et al.
Veröffentlicht: (2026)
Integrating Pause Information with Word Embeddings in Language Models for Alzheimer's Disease Detection from Spontaneous Speech
von: Pu, Yu, et al.
Veröffentlicht: (2025)
von: Pu, Yu, et al.
Veröffentlicht: (2025)
Two-Stage Adaptation for Non-Normative Speech Recognition: Revisiting Speaker-Independent Initialization for Personalization
von: Jiang, Shan, et al.
Veröffentlicht: (2026)
von: Jiang, Shan, et al.
Veröffentlicht: (2026)
A Two-Stage Hierarchical Deep Filtering Framework for Real-Time Speech Enhancement
von: Lu, Shenghui, et al.
Veröffentlicht: (2025)
von: Lu, Shenghui, et al.
Veröffentlicht: (2025)
HAFFormer: A Hierarchical Attention-Free Framework for Alzheimer's Disease Detection From Spontaneous Speech
von: Dong, Zhongren, et al.
Veröffentlicht: (2024)
von: Dong, Zhongren, et al.
Veröffentlicht: (2024)
Speech as a Biomarker for Disease Detection
von: Botelho, Catarina, et al.
Veröffentlicht: (2024)
von: Botelho, Catarina, et al.
Veröffentlicht: (2024)
Speech-Mamba: Long-Context Speech Recognition with Selective State Spaces Models
von: Gao, Xiaoxue, et al.
Veröffentlicht: (2024)
von: Gao, Xiaoxue, et al.
Veröffentlicht: (2024)
InconVAD: A Two-Stage Dual-Tower Framework for Multimodal Emotion Inconsistency Detection
von: Li, Zongyi, et al.
Veröffentlicht: (2025)
von: Li, Zongyi, et al.
Veröffentlicht: (2025)
Reverse-Speech-Finder: A Neural Network Backtracking Architecture for Generating Alzheimer's Disease Speech Samples and Improving Diagnosis Performance
von: Li, Victor OK, et al.
Veröffentlicht: (2025)
von: Li, Victor OK, et al.
Veröffentlicht: (2025)
Towards Explicit Acoustic Evidence Perception in Audio LLMs for Speech Deepfake Detection
von: Guo, Xiaoxuan, et al.
Veröffentlicht: (2026)
von: Guo, Xiaoxuan, et al.
Veröffentlicht: (2026)
AffectCodec: Emotion-Preserving Neural Speech Codec for Expressive Speech Modeling
von: Shi, Jiacheng, et al.
Veröffentlicht: (2026)
von: Shi, Jiacheng, et al.
Veröffentlicht: (2026)
A Two-Stage Framework in Cross-Spectrum Domain for Real-Time Speech Enhancement
von: Zhang, Yuewei, et al.
Veröffentlicht: (2024)
von: Zhang, Yuewei, et al.
Veröffentlicht: (2024)
EMO-TTA: Improving Test-Time Adaptation of Audio-Language Models for Speech Emotion Recognition
von: Shi, Jiacheng, et al.
Veröffentlicht: (2025)
von: Shi, Jiacheng, et al.
Veröffentlicht: (2025)
BridgeCode: A Dual Speech Representation Paradigm for Autoregressive Zero-Shot Text-to-Speech Synthesis
von: Xing, Jingyuan, et al.
Veröffentlicht: (2025)
von: Xing, Jingyuan, et al.
Veröffentlicht: (2025)
Fake Speech Wild: Detecting Deepfake Speech on Social Media Platform
von: Xie, Yuankun, et al.
Veröffentlicht: (2025)
von: Xie, Yuankun, et al.
Veröffentlicht: (2025)
Emphasis Rendering for Conversational Text-to-Speech with Multi-modal Multi-scale Context Modeling
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
DementiaBank-Emotion: A Multi-Rater Emotion Annotation Corpus for Alzheimer's Disease Speech (Version 1.0)
von: Jeong, Cheonkam, et al.
Veröffentlicht: (2026)
von: Jeong, Cheonkam, et al.
Veröffentlicht: (2026)
DAST: A Dual-Stream Voice Anonymization Attacker with Staged Training
von: Arefeen, Ridwan, et al.
Veröffentlicht: (2026)
von: Arefeen, Ridwan, et al.
Veröffentlicht: (2026)
Fine-Grained Frame Modeling in Multi-head Self-Attention for Speech Deepfake Detection
von: Phuong, Tuan Dat, et al.
Veröffentlicht: (2026)
von: Phuong, Tuan Dat, et al.
Veröffentlicht: (2026)
EnvTriCascade: An Environment-Aware Tri-Stage Cascaded Framework for ESDD2 2026 Challenge
von: Huang, Hengyan, et al.
Veröffentlicht: (2026)
von: Huang, Hengyan, et al.
Veröffentlicht: (2026)
Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition
von: Bai, Ye, et al.
Veröffentlicht: (2024)
von: Bai, Ye, et al.
Veröffentlicht: (2024)
Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection
von: Amiri, Mahdi, et al.
Veröffentlicht: (2025)
von: Amiri, Mahdi, et al.
Veröffentlicht: (2025)
A Novel Automatic Framework for Speaker Drift Detection in Synthesized Speech
von: Huang, Jia-Hong, et al.
Veröffentlicht: (2026)
von: Huang, Jia-Hong, et al.
Veröffentlicht: (2026)
Dynamic Fusion Multimodal Network for SpeechWellness Detection
von: Sun, Wenqiang, et al.
Veröffentlicht: (2025)
von: Sun, Wenqiang, et al.
Veröffentlicht: (2025)
Dual-Branch Knowledge Distillation for Noise-Robust Synthetic Speech Detection
von: Fan, Cunhang, et al.
Veröffentlicht: (2023)
von: Fan, Cunhang, et al.
Veröffentlicht: (2023)
Evaluating Parkinson's Disease Detection in Anonymized Speech: A Performance and Acoustic Analysis
von: Franzreb, Carlos, et al.
Veröffentlicht: (2026)
von: Franzreb, Carlos, et al.
Veröffentlicht: (2026)
Dual Data Scaling for Robust Two-Stage User-Defined Keyword Spotting
von: Ai, Zhiqi, et al.
Veröffentlicht: (2025)
von: Ai, Zhiqi, et al.
Veröffentlicht: (2025)
MultiAPI Spoof: A Multi-API Dataset and Local-Attention Network for Speech Anti-spoofing Detection
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
von: Zhang, Xueping, et al.
Veröffentlicht: (2025)
DisCo-Speech: Controllable Zero-Shot Speech Generation with A Disentangled Speech Codec
von: Li, Tao, et al.
Veröffentlicht: (2025)
von: Li, Tao, et al.
Veröffentlicht: (2025)
MoTAS: MoE-Guided Feature Selection from TTS-Augmented Speech for Enhanced Multimodal Alzheimer's Early Screening
von: Shao, Yongqi, et al.
Veröffentlicht: (2025)
von: Shao, Yongqi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Leveraging Multimodal Methods and Spontaneous Speech for Alzheimer's Disease Identification
von: Gao, Yifan, et al.
Veröffentlicht: (2024) -
Leveraging Large Language Models for Spontaneous Speech-Based Suicide Risk Detection
von: Gao, Yifan, et al.
Veröffentlicht: (2025) -
Leveraging Prompt Learning and Pause Encoding for Alzheimer's Disease Detection
von: Liu, Yin-Long, et al.
Veröffentlicht: (2024) -
Exploring Gender Bias in Alzheimer's Disease Detection: Insights from Mandarin and Greek Speech Perception
von: He, Liu, et al.
Veröffentlicht: (2025) -
Two-Stage Acoustic Adaptation with Gated Cross-Attention Adapters for LLM-Based Multi-Talker Speech Recognition
von: Shi, Hao, et al.
Veröffentlicht: (2026)