Rene: A Pre-trained Multi-modal Architecture for Auscultation of Respiratory Diseases
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Pengfei, Zheng, Zhihang, Zhang, Shichen, Yang, Minghao, Tang, Shaojun |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multi Modal Information Fusion of Acoustic and Linguistic Data for Decoding Dairy Cow Vocalizations in Animal Welfare Assessment
por: Jobarteh, Bubacarr, et al.
Publicado: (2024)
por: Jobarteh, Bubacarr, et al.
Publicado: (2024)
Anonymising Elderly and Pathological Speech: Voice Conversion Using DDSP and Query-by-Example
por: Ghosh, Suhita, et al.
Publicado: (2024)
por: Ghosh, Suhita, et al.
Publicado: (2024)
Temporal Feature Learning in Weakly Labelled Bioacoustic Cetacean Datasets via a Variational Autoencoder and Temporal Convolutional Network: An Interdisciplinary Approach
por: Fonollosa, Laia Garrobé, et al.
Publicado: (2024)
por: Fonollosa, Laia Garrobé, et al.
Publicado: (2024)
Computational bioacoustics with deep learning: a review and roadmap
por: Stowell, Dan
Publicado: (2021)
por: Stowell, Dan
Publicado: (2021)
Adaptive Representations of Sound for Automatic Insect Recognition
por: Faiß, Marius, et al.
Publicado: (2023)
por: Faiß, Marius, et al.
Publicado: (2023)
Learning to detect an animal sound from five examples
por: Nolasco, Inês, et al.
Publicado: (2023)
por: Nolasco, Inês, et al.
Publicado: (2023)
Fish Tracking, Counting, and Behaviour Analysis in Digital Aquaculture: A Comprehensive Survey
por: Cui, Meng, et al.
Publicado: (2024)
por: Cui, Meng, et al.
Publicado: (2024)
animal2vec and MeerKAT: A self-supervised transformer for rare-event raw audio input and a large-scale reference dataset for bioacoustics
por: Schäfer-Zimmermann, Julian C., et al.
Publicado: (2024)
por: Schäfer-Zimmermann, Julian C., et al.
Publicado: (2024)
Automatic detection of Mild Cognitive Impairment using high-dimensional acoustic features in spontaneous speech
por: Zhang, Cong, et al.
Publicado: (2024)
por: Zhang, Cong, et al.
Publicado: (2024)
Prosody of speech production in latent post-stroke aphasia
por: Zhang, Cong, et al.
Publicado: (2024)
por: Zhang, Cong, et al.
Publicado: (2024)
Automated Measurement of Geniohyoid Muscle Thickness During Speech Using Deep Learning and Ultrasound
por: Myrgyyassov, Alisher, et al.
Publicado: (2026)
por: Myrgyyassov, Alisher, et al.
Publicado: (2026)
WhaleVAD-BPN: Improving Baleen Whale Call Detection with Boundary Proposal Networks and Post-processing Optimisation
por: Geldenhuys, Christiaan M., et al.
Publicado: (2025)
por: Geldenhuys, Christiaan M., et al.
Publicado: (2025)
Cochlear Wave Propagation and Dynamics in the Human Base and Apex: Model-Based Estimates from Noninvasive Measurements
por: Alkhairy, Samiya A
Publicado: (2024)
por: Alkhairy, Samiya A
Publicado: (2024)
A Classification Benchmark for Artificial Intelligence Detection of Laryngeal Cancer from Patient Voice
por: Paterson, Mary, et al.
Publicado: (2024)
por: Paterson, Mary, et al.
Publicado: (2024)
Which Augmentation Should I Use? An Empirical Investigation of Augmentations for Self-Supervised Phonocardiogram Representation Learning
por: Ballas, Aristotelis, et al.
Publicado: (2023)
por: Ballas, Aristotelis, et al.
Publicado: (2023)
Screening method for early dementia using sound objects as voice biomarkers
por: Pluta, Adam, et al.
Publicado: (2024)
por: Pluta, Adam, et al.
Publicado: (2024)
Foundation Models for Bioacoustics -- a Comparative Review
por: Schwinger, Raphael, et al.
Publicado: (2025)
por: Schwinger, Raphael, et al.
Publicado: (2025)
Learning to rumble: Automated elephant call classification, detection and endpointing using deep architectures
por: Geldenhuys, Christiaan M., et al.
Publicado: (2024)
por: Geldenhuys, Christiaan M., et al.
Publicado: (2024)
From Birdsong to Rumbles: Classifying Elephant Calls with Out-of-Species Embeddings
por: Geldenhuys, Christiaan M., et al.
Publicado: (2026)
por: Geldenhuys, Christiaan M., et al.
Publicado: (2026)
All Thresholds Barred: Direct Estimation of Call Density in Bioacoustic Data
por: Navine, Amanda K., et al.
Publicado: (2024)
por: Navine, Amanda K., et al.
Publicado: (2024)
Atrial Fibrillation Detection System via Acoustic Sensing for Mobile Phones
por: Liu, Xuanyu, et al.
Publicado: (2024)
por: Liu, Xuanyu, et al.
Publicado: (2024)
Patient-Level Multimodal Question Answering from Multi-Site Auscultation Recordings
por: Wu, Fan, et al.
Publicado: (2026)
por: Wu, Fan, et al.
Publicado: (2026)
Semi-supervised classification of bird vocalizations
por: Hexeberg, Simen, et al.
Publicado: (2025)
por: Hexeberg, Simen, et al.
Publicado: (2025)
A Generalist Audio Foundation Model for Comprehensive Body Sound Auscultation
por: Wang, Pingjie, et al.
Publicado: (2024)
por: Wang, Pingjie, et al.
Publicado: (2024)
Intelligent Cardiac Auscultation for Murmur Detection via Parallel-Attentive Models with Uncertainty Estimation
por: Zhang, Zixing, et al.
Publicado: (2024)
por: Zhang, Zixing, et al.
Publicado: (2024)
Prior-agnostic Multi-scale Contrastive Text-Audio Pre-training for Parallelized TTS Frontend Modeling
por: Wang, Quanxiu, et al.
Publicado: (2024)
por: Wang, Quanxiu, et al.
Publicado: (2024)
MAT-SED: A Masked Audio Transformer with Masked-Reconstruction Based Pre-training for Sound Event Detection
por: Cai, Pengfei, et al.
Publicado: (2024)
por: Cai, Pengfei, et al.
Publicado: (2024)
Active Learning with Task Adaptation Pre-training for Speech Emotion Recognition
por: Li, Dongyuan, et al.
Publicado: (2024)
por: Li, Dongyuan, et al.
Publicado: (2024)
Improving Respiratory Sound Classification with Architecture-Agnostic Knowledge Distillation from Ensembles
por: Toikkanen, Miika, et al.
Publicado: (2025)
por: Toikkanen, Miika, et al.
Publicado: (2025)
Mozart's Touch: A Lightweight Multi-modal Music Generation Framework Based on Pre-Trained Large Models
por: Li, Jiajun, et al.
Publicado: (2024)
por: Li, Jiajun, et al.
Publicado: (2024)
Review of Cetacean's click detection algorithms
por: Gracic, Mak, et al.
Publicado: (2024)
por: Gracic, Mak, et al.
Publicado: (2024)
Temporal Adaptation of Pre-trained Foundation Models for Music Structure Analysis
por: Zhang, Yixiao, et al.
Publicado: (2025)
por: Zhang, Yixiao, et al.
Publicado: (2025)
Adaptive Differential Denoising for Respiratory Sounds Classification
por: Dong, Gaoyang, et al.
Publicado: (2025)
por: Dong, Gaoyang, et al.
Publicado: (2025)
SongGLM: Lyric-to-Melody Generation with 2D Alignment Encoding and Multi-Task Pre-Training
por: Yu, Jiaxing, et al.
Publicado: (2024)
por: Yu, Jiaxing, et al.
Publicado: (2024)
DSFlow: Dual Supervision and Step-Aware Architecture for One-Step Flow Matching Speech Synthesis
por: Lin, Bin, et al.
Publicado: (2026)
por: Lin, Bin, et al.
Publicado: (2026)
SONAR: Self-Distilled Continual Pre-training for Domain Adaptive Audio Representation
por: Zhang, Yizhou, et al.
Publicado: (2025)
por: Zhang, Yizhou, et al.
Publicado: (2025)
Efficient Speech Enhancement via Embeddings from Pre-trained Generative Audioencoders
por: Sun, Xingwei, et al.
Publicado: (2025)
por: Sun, Xingwei, et al.
Publicado: (2025)
A multimodal LLM for the non-invasive decoding of spoken text from brain recordings
por: Hmamouche, Youssef, et al.
Publicado: (2024)
por: Hmamouche, Youssef, et al.
Publicado: (2024)
Sound training platform applied to astronomy
por: Lucero, Natasha Bertaina, et al.
Publicado: (2024)
por: Lucero, Natasha Bertaina, et al.
Publicado: (2024)
Disambiguation of Chinese Polyphones in an End-to-End Framework with Semantic Features Extracted by Pre-trained BERT
por: Dai, Dongyang, et al.
Publicado: (2025)
por: Dai, Dongyang, et al.
Publicado: (2025)
Ejemplares similares
-
Multi Modal Information Fusion of Acoustic and Linguistic Data for Decoding Dairy Cow Vocalizations in Animal Welfare Assessment
por: Jobarteh, Bubacarr, et al.
Publicado: (2024) -
Anonymising Elderly and Pathological Speech: Voice Conversion Using DDSP and Query-by-Example
por: Ghosh, Suhita, et al.
Publicado: (2024) -
Temporal Feature Learning in Weakly Labelled Bioacoustic Cetacean Datasets via a Variational Autoencoder and Temporal Convolutional Network: An Interdisciplinary Approach
por: Fonollosa, Laia Garrobé, et al.
Publicado: (2024) -
Computational bioacoustics with deep learning: a review and roadmap
por: Stowell, Dan
Publicado: (2021) -
Adaptive Representations of Sound for Automatic Insect Recognition
por: Faiß, Marius, et al.
Publicado: (2023)