Localization using Angle-of-Arrival Triangulation
Fuente:
arXiv
Guardado en:
| Autor principal: | Agrawal, Amod K. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MaskClip: Detachable Clip-on Piezoelectric Sensing of Mask Surface Vibrations for Real-time Noise-Robust Speech Input
por: Hiraki, Hirotaka, et al.
Publicado: (2025)
por: Hiraki, Hirotaka, et al.
Publicado: (2025)
Fine-Tuning Large Audio-Language Models with LoRA for Precise Temporal Localization of Prolonged Exposure Therapy Elements
por: BN, Suhas, et al.
Publicado: (2025)
por: BN, Suhas, et al.
Publicado: (2025)
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
por: Kim, Minu, et al.
Publicado: (2025)
por: Kim, Minu, et al.
Publicado: (2025)
Thaka at KSAA-2026 Task 2: Regularized Fine-Tuning for Arabic Speech Diacritization
por: Alamr, Meshal, et al.
Publicado: (2026)
por: Alamr, Meshal, et al.
Publicado: (2026)
Predicting Upcoming Stuttering Events from Three-Second Audio: Stratified Evaluation Reveals Severity-Selective Precursors, and the Model Deploys Fully On-Device
por: Kozak, Nazar
Publicado: (2026)
por: Kozak, Nazar
Publicado: (2026)
Quantum-Enhanced Analysis and Grading of Vocal Performance
por: Agarwal, Rohan
Publicado: (2025)
por: Agarwal, Rohan
Publicado: (2025)
3GPP NR V2X Mode 2d: Analysis of Distributed Scheduling for Groupcast using ns-3 5G LENA Simulator
por: Fehrenbach, Thomas, et al.
Publicado: (2025)
por: Fehrenbach, Thomas, et al.
Publicado: (2025)
Improving Cross-Lingual Phonetic Representation of Low-Resource Languages Through Language Similarity Analysis
por: Kim, Minu, et al.
Publicado: (2025)
por: Kim, Minu, et al.
Publicado: (2025)
Delayed Fusion: Integrating Large Language Models into First-Pass Decoding in End-to-end Speech Recognition
por: Hori, Takaaki, et al.
Publicado: (2025)
por: Hori, Takaaki, et al.
Publicado: (2025)
An End-to-End Approach for Korean Wakeword Systems with Speaker Authentication
por: Seo, Geonwoo
Publicado: (2025)
por: Seo, Geonwoo
Publicado: (2025)
Adaptive Edge-Cloud Inference for Speech-to-Action Systems Using ASR and Large Language Models
por: Torkamani, Mohammad Jalili, et al.
Publicado: (2025)
por: Torkamani, Mohammad Jalili, et al.
Publicado: (2025)
Emotional Voice Messages (EMOVOME) database: emotion recognition in spontaneous voice messages
por: Zaragozá, Lucía Gómez, et al.
Publicado: (2024)
por: Zaragozá, Lucía Gómez, et al.
Publicado: (2024)
Sink or SWIM: Tackling Real-Time ASR at Scale
por: Bruzzone, Federico, et al.
Publicado: (2026)
por: Bruzzone, Federico, et al.
Publicado: (2026)
The evolution of inharmonicity and noisiness in contemporary popular music
por: Deruty, Emmanuel, et al.
Publicado: (2024)
por: Deruty, Emmanuel, et al.
Publicado: (2024)
Splitformer: An improved early-exit architecture for automatic speech recognition on edge devices
por: Lasbordes, Maxence, et al.
Publicado: (2025)
por: Lasbordes, Maxence, et al.
Publicado: (2025)
What Would GPT Click: Practical Effects of Human-AI Behavioral Misalignment and the Cost of Synthetic Participants in User Experience
por: Kuric, Eduard, et al.
Publicado: (2026)
por: Kuric, Eduard, et al.
Publicado: (2026)
The OCON model: an old but green solution for distributable supervised classification for acoustic monitoring in smart cities
por: Giacomelli, Stefano, et al.
Publicado: (2024)
por: Giacomelli, Stefano, et al.
Publicado: (2024)
STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts
por: Opria, Joshua
Publicado: (2026)
por: Opria, Joshua
Publicado: (2026)
Hidden Echoes Survive Training in Audio To Audio Generative Instrument Models
por: Tralie, Christopher J., et al.
Publicado: (2024)
por: Tralie, Christopher J., et al.
Publicado: (2024)
Machine Learning Framework for Audio-Based Content Evaluation using MFCC, Chroma, Spectral Contrast, and Temporal Feature Engineering
por: Aristorenas, Aris J.
Publicado: (2024)
por: Aristorenas, Aris J.
Publicado: (2024)
Revisiting SSL for sound event detection: complementary fusion and adaptive post-processing
por: Cui, Hanfang, et al.
Publicado: (2025)
por: Cui, Hanfang, et al.
Publicado: (2025)
Joint Estimation of Piano Dynamics and Metrical Structure with a Multi-task Multi-Scale Network
por: He, Zhanhong, et al.
Publicado: (2025)
por: He, Zhanhong, et al.
Publicado: (2025)
GIST: Multimodal Knowledge Extraction and Spatial Grounding via Intelligent Semantic Topology
por: Agrawal, Shivendra, et al.
Publicado: (2026)
por: Agrawal, Shivendra, et al.
Publicado: (2026)
Enhanced DareFightingICE Competitions: Sound Design and AI Competitions
por: Khan, Ibrahim, et al.
Publicado: (2024)
por: Khan, Ibrahim, et al.
Publicado: (2024)
Impact of Phonetics on Speaker Identity in Adversarial Voice Attack
por: Dar, Daniyal Kabir, et al.
Publicado: (2025)
por: Dar, Daniyal Kabir, et al.
Publicado: (2025)
WhisperMask: A Noise Suppressive Mask-Type Microphone for Whisper Speech
por: Hiraki, Hirotaka, et al.
Publicado: (2024)
por: Hiraki, Hirotaka, et al.
Publicado: (2024)
Whisphone: Whispering Input Earbuds
por: Fukumoto, Masaaki
Publicado: (2025)
por: Fukumoto, Masaaki
Publicado: (2025)
Operational Latent Spaces
por: Hawley, Scott H., et al.
Publicado: (2024)
por: Hawley, Scott H., et al.
Publicado: (2024)
Modeling L1 Influence on L2 Pronunciation: An MFCC-Based Framework for Explainable Machine Learning and Pedagogical Feedback
por: Jahanbin, Peyman
Publicado: (2025)
por: Jahanbin, Peyman
Publicado: (2025)
emg2qwerty: A Large Dataset with Baselines for Touch Typing using Surface Electromyography
por: Sivakumar, Viswanath, et al.
Publicado: (2024)
por: Sivakumar, Viswanath, et al.
Publicado: (2024)
Improving Speech Recognition Accuracy Using Custom Language Models with the Vosk Toolkit
por: Soni, Aniket Abhishek
Publicado: (2025)
por: Soni, Aniket Abhishek
Publicado: (2025)
Deep Feed-Forward Neural Network for Bangla Isolated Speech Recognition
por: Bhadra, Dipayan, et al.
Publicado: (2025)
por: Bhadra, Dipayan, et al.
Publicado: (2025)
Distilled HuBERT for Mobile Speech Emotion Recognition: A Cross-Corpus Validation Study
por: Ismail, Saifelden M.
Publicado: (2025)
por: Ismail, Saifelden M.
Publicado: (2025)
Fine-tuning Pre-trained Audio Models for COVID-19 Detection: A Technical Report
por: de Brito, Daniel Oliveira, et al.
Publicado: (2025)
por: de Brito, Daniel Oliveira, et al.
Publicado: (2025)
Audio-based Kinship Verification Using Age Domain Conversion
por: Sun, Qiyang, et al.
Publicado: (2024)
por: Sun, Qiyang, et al.
Publicado: (2024)
Prevailing Research Areas for Music AI in the Era of Foundation Models
por: Wei, Megan, et al.
Publicado: (2024)
por: Wei, Megan, et al.
Publicado: (2024)
Silent Impact: Tracking Tennis Shots from the Passive Arm
por: Park, Junyong, et al.
Publicado: (2025)
por: Park, Junyong, et al.
Publicado: (2025)
OBHS: An Optimized Block Huffman Scheme for Real-Time Audio Compression
por: Mahfi, Muntahi Safwan, et al.
Publicado: (2025)
por: Mahfi, Muntahi Safwan, et al.
Publicado: (2025)
MAC-Gaze: Motion-Aware Continual Calibration for Mobile Gaze Tracking
por: Lei, Yaxiong, et al.
Publicado: (2025)
por: Lei, Yaxiong, et al.
Publicado: (2025)
Passive Underwater Acoustic Signal Separation based on Feature Decoupling Dual-path Network
por: Liu, Yucheng, et al.
Publicado: (2025)
por: Liu, Yucheng, et al.
Publicado: (2025)
Ejemplares similares
-
MaskClip: Detachable Clip-on Piezoelectric Sensing of Mask Surface Vibrations for Real-time Noise-Robust Speech Input
por: Hiraki, Hirotaka, et al.
Publicado: (2025) -
Fine-Tuning Large Audio-Language Models with LoRA for Precise Temporal Localization of Prolonged Exposure Therapy Elements
por: BN, Suhas, et al.
Publicado: (2025) -
ParaNoise-SV: Integrated Approach for Noise-Robust Speaker Verification with Parallel Joint Learning of Speech Enhancement and Noise Extraction
por: Kim, Minu, et al.
Publicado: (2025) -
Thaka at KSAA-2026 Task 2: Regularized Fine-Tuning for Arabic Speech Diacritization
por: Alamr, Meshal, et al.
Publicado: (2026) -
Predicting Upcoming Stuttering Events from Three-Second Audio: Stratified Evaluation Reveals Severity-Selective Precursors, and the Model Deploys Fully On-Device
por: Kozak, Nazar
Publicado: (2026)