Scaling and Distilling Transformer Models for sEMG
Fuente:
arXiv
Saved in:
| Main Authors: | Mehlman, Nicholas, Gagnon-Audet, Jean-Christophe, Shvartsman, Michael, Niu, Kelvin, Miller, Alexander H., Sodhani, Shagun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A 1000-hour EEG-EMG-audio dataset of Japanese speech production
by: Sato, Motoshige, et al.
Published: (2026)
by: Sato, Motoshige, et al.
Published: (2026)
Poster: Recognizing Hidden-in-the-Ear Private Key for Reliable Silent Speech Interface Using Multi-Task Learning
by: Dong, Xuefu, et al.
Published: (2025)
by: Dong, Xuefu, et al.
Published: (2025)
Cluster-to-Predict Affect Contours from Speech
by: Kuşçu, Gökhan, et al.
Published: (2024)
by: Kuşçu, Gökhan, et al.
Published: (2024)
Silent Speech Sentence Recognition with Six-Axis Accelerometers using Conformer and CTC Algorithm
by: Xie, Yudong, et al.
Published: (2025)
by: Xie, Yudong, et al.
Published: (2025)
Toward using Speech to Sense Student Emotion in Remote Learning Environments
by: Vyas, Sargam, et al.
Published: (2026)
by: Vyas, Sargam, et al.
Published: (2026)
Collecting Prosody in the Wild: A Content-Controlled, Privacy-First Smartphone Protocol and Empirical Evaluation
by: Koch, Timo K., et al.
Published: (2026)
by: Koch, Timo K., et al.
Published: (2026)
Timbre-Aware LLM-based Direct Speech-to-Speech Translation Extendable to Multiple Language Pairs
by: Arya, Lalaram, et al.
Published: (2026)
by: Arya, Lalaram, et al.
Published: (2026)
UltrasonicSpheres: Localized, Multi-Channel Sound Spheres Using Off-the-Shelf Speakers and Earables
by: Küttner, Michael, et al.
Published: (2025)
by: Küttner, Michael, et al.
Published: (2025)
Towards Temporally Explainable Dysarthric Speech Clarity Assessment
by: Park, Seohyun, et al.
Published: (2025)
by: Park, Seohyun, et al.
Published: (2025)
USpeech: Ultrasound-Enhanced Speech with Minimal Human Effort via Cross-Modal Synthesis
by: Yu, Luca Jiang-Tao, et al.
Published: (2024)
by: Yu, Luca Jiang-Tao, et al.
Published: (2024)
Early Detection of Furniture-Infesting Wood-Boring Beetles Using CNN-LSTM Networks and MFCC-Based Acoustic Features
by: Manukalpa, J. M. Chan Sri, et al.
Published: (2025)
by: Manukalpa, J. M. Chan Sri, et al.
Published: (2025)
Personalized Speech Emotion Recognition in Human-Robot Interaction using Vision Transformers
by: Mishra, Ruchik, et al.
Published: (2024)
by: Mishra, Ruchik, et al.
Published: (2024)
Open Your Ears and Take a Look: A State-of-the-Art Report on the Integration of Sonification and Visualization
by: Enge, Kajetan, et al.
Published: (2024)
by: Enge, Kajetan, et al.
Published: (2024)
Inferring trust in recommendation systems from brain, behavioural, and physiological data
by: Cheung, Vincent K. M., et al.
Published: (2025)
by: Cheung, Vincent K. M., et al.
Published: (2025)
Real-Time Auralization for First-Person Vocal Interaction in Immersive Virtual Environments
by: Flores-Vargas, Mauricio, et al.
Published: (2025)
by: Flores-Vargas, Mauricio, et al.
Published: (2025)
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
by: Khanday, Owais Mujtaba, et al.
Published: (2025)
by: Khanday, Owais Mujtaba, et al.
Published: (2025)
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
by: Khanday, Owais Mujtaba, et al.
Published: (2025)
by: Khanday, Owais Mujtaba, et al.
Published: (2025)
Sound2Hap: Learning Audio-to-Vibrotactile Haptic Generation from Human Ratings
by: Li, Yinan, et al.
Published: (2026)
by: Li, Yinan, et al.
Published: (2026)
A Mapping Strategy for Interacting with Latent Audio Synthesis Using Artistic Materials
by: Zheng, Shuoyang, et al.
Published: (2024)
by: Zheng, Shuoyang, et al.
Published: (2024)
Enhancing DMI Interactions by Integrating Haptic Feedback for Intricate Vibrato Technique
by: Piao, Ziyue, et al.
Published: (2024)
by: Piao, Ziyue, et al.
Published: (2024)
Enhancing the NAO: Extending Capabilities of Legacy Robots for Long-Term Research
by: Wilson, Austin, et al.
Published: (2025)
by: Wilson, Austin, et al.
Published: (2025)
A cross-talk robust multichannel VAD model for multiparty agent interactions trained using synthetic re-recordings
by: Han, Hyewon, et al.
Published: (2024)
by: Han, Hyewon, et al.
Published: (2024)
Interactive Sonification for Health and Energy using ChucK and Unity
by: Zhao, Yichun, et al.
Published: (2024)
by: Zhao, Yichun, et al.
Published: (2024)
A Near-Real-Time Processing Ego Speech Filtering Pipeline Designed for Speech Interruption During Human-Robot Interaction
by: Li, Yue, et al.
Published: (2024)
by: Li, Yue, et al.
Published: (2024)
Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
by: Blum'e, Ashlae
Published: (2025)
by: Blum'e, Ashlae
Published: (2025)
Interfacing with history: Curating with audio augmented objects
by: Cliffe, Laurence
Published: (2024)
by: Cliffe, Laurence
Published: (2024)
Transhuman Ansambl - Voice Beyond Language
by: Ivsic, Lucija, et al.
Published: (2024)
by: Ivsic, Lucija, et al.
Published: (2024)
How Private is Low-Frequency Speech Audio in the Wild? An Analysis of Verbal Intelligibility by Humans and Machines
by: Liu, Ailin, et al.
Published: (2024)
by: Liu, Ailin, et al.
Published: (2024)
Cervical Auscultation Machine Learning for Dysphagia Assessment
by: Chia, An An, et al.
Published: (2024)
by: Chia, An An, et al.
Published: (2024)
ExSampling: a system for the real-time ensemble performance of field-recorded environmental sounds
by: Kobayashi, Atsuya, et al.
Published: (2020)
by: Kobayashi, Atsuya, et al.
Published: (2020)
Adapting Whisper for Lightweight and Efficient Automatic Speech Recognition of Children for On-device Edge Applications
by: Dutta, Satwik, et al.
Published: (2025)
by: Dutta, Satwik, et al.
Published: (2025)
Real-time Generation of Various Types of Nodding for Avatar Attentive Listening System
by: Kato, Kazushi, et al.
Published: (2025)
by: Kato, Kazushi, et al.
Published: (2025)
BioSonix: Can Physics-Based Sonification Perceptualize Tissue Deformations From Tool Interactions?
by: Ruozzi, Veronica, et al.
Published: (2025)
by: Ruozzi, Veronica, et al.
Published: (2025)
Springboard, Roadblock or "Crutch"?: How Transgender Users Leverage Voice Changers for Gender Presentation in Social Virtual Reality
by: Povinelli, Kassie, et al.
Published: (2024)
by: Povinelli, Kassie, et al.
Published: (2024)
Evolving Performance Practices in Beethoven's Cello Sonatas: Tempo, Portamento, and Historical Interpretation of the First Movements
by: Sole, Ignasi
Published: (2025)
by: Sole, Ignasi
Published: (2025)
SCDiar: a streaming diarization system based on speaker change detection and speech recognition
by: Zheng, Naijun, et al.
Published: (2025)
by: Zheng, Naijun, et al.
Published: (2025)
Teach Me How to ImproVISe: Co-Designing an Augmented Piano Training System for Improvisation
by: Deja, Jordan Aiko, et al.
Published: (2024)
by: Deja, Jordan Aiko, et al.
Published: (2024)
Open vocabulary keyword spotting through transfer learning from speech synthesis
by: V, Kesavaraj, et al.
Published: (2024)
by: V, Kesavaraj, et al.
Published: (2024)
The effect of self-motion and room familiarity on sound source localization in virtual environments
by: Isserstedt, Niklas, et al.
Published: (2024)
by: Isserstedt, Niklas, et al.
Published: (2024)
NeckCare: Preventing Tech Neck using Hearable-based Multimodal Sensing
by: Chhaglani, Bhawana, et al.
Published: (2024)
by: Chhaglani, Bhawana, et al.
Published: (2024)
Similar Items
-
A 1000-hour EEG-EMG-audio dataset of Japanese speech production
by: Sato, Motoshige, et al.
Published: (2026) -
Poster: Recognizing Hidden-in-the-Ear Private Key for Reliable Silent Speech Interface Using Multi-Task Learning
by: Dong, Xuefu, et al.
Published: (2025) -
Cluster-to-Predict Affect Contours from Speech
by: Kuşçu, Gökhan, et al.
Published: (2024) -
Silent Speech Sentence Recognition with Six-Axis Accelerometers using Conformer and CTC Algorithm
by: Xie, Yudong, et al.
Published: (2025) -
Toward using Speech to Sense Student Emotion in Remote Learning Environments
by: Vyas, Sargam, et al.
Published: (2026)