MathSpeech: Leveraging Small LMs for Accurate Conversion in Mathematical Speech-to-Formula
Fuente:
arXiv
Saved in:
| Main Authors: | Hyeon, Sieun, Jung, Kyudan, Won, Jaehee, Kim, Nam-Joon, Ryu, Hyun Gon, Lee, Hyuk-Jae, Do, Jaeyoung |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MathReader : Text-to-Speech for Mathematical Documents
by: Hyeon, Sieun, et al.
Published: (2025)
by: Hyeon, Sieun, et al.
Published: (2025)
MathBridge: A Large Corpus Dataset for Translating Spoken Mathematical Expressions into $LaTeX$ Formulas for Improved Readability
by: Jung, Kyudan, et al.
Published: (2024)
by: Jung, Kyudan, et al.
Published: (2024)
Enhancing ASR Performance through OCR Word Frequency Analysis: Theoretical Foundations
by: Jung, Kyudan, et al.
Published: (2024)
by: Jung, Kyudan, et al.
Published: (2024)
Is Retraining-Free Enough? The Necessity of Router Calibration for Efficient MoE Compression
by: Hyeon, Sieun, et al.
Published: (2026)
by: Hyeon, Sieun, et al.
Published: (2026)
TeXBLEU: Automatic Metric for Evaluate LaTeX Format
by: Jung, Kyudan, et al.
Published: (2024)
by: Jung, Kyudan, et al.
Published: (2024)
VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models
by: Kim, Woojin, et al.
Published: (2026)
by: Kim, Woojin, et al.
Published: (2026)
MATA: Multi-Agent Framework for Reliable and Flexible Table Question Answering
by: Hyeon, Sieun, et al.
Published: (2026)
by: Hyeon, Sieun, et al.
Published: (2026)
SLICE: Speech Enhancement via Layer-wise Injection of Conditioning Embeddings
by: Moon, Seokhoon, et al.
Published: (2026)
by: Moon, Seokhoon, et al.
Published: (2026)
Automatic Speech Recognition (ASR) for the Diagnosis of pronunciation of Speech Sound Disorders in Korean children
by: Ahn, Taekyung, et al.
Published: (2024)
by: Ahn, Taekyung, et al.
Published: (2024)
Latent-Level Enhancement with Flow Matching for Robust Automatic Speech Recognition
by: Yang, Da-Hee, et al.
Published: (2026)
by: Yang, Da-Hee, et al.
Published: (2026)
Exploring and Leveraging Class Vectors for Classifier Editing
by: Kim, Jaeik, et al.
Published: (2025)
by: Kim, Jaeik, et al.
Published: (2025)
SNAP: Speaker Nulling for Artifact Projection in Speech Deepfake Detection
by: Jung, Kyudan, et al.
Published: (2026)
by: Jung, Kyudan, et al.
Published: (2026)
Whisfusion: Parallel ASR Decoding via a Diffusion Transformer
by: Kwon, Taeyoun, et al.
Published: (2025)
by: Kwon, Taeyoun, et al.
Published: (2025)
Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems
by: Wei, Chengwei, et al.
Published: (2025)
by: Wei, Chengwei, et al.
Published: (2025)
Sommelier: Scalable Open Multi-turn Audio Pre-processing for Full-duplex Speech Language Models
by: Jung, Kyudan, et al.
Published: (2026)
by: Jung, Kyudan, et al.
Published: (2026)
Two Dimensions of Financial Openness: Gross and Net Effects on Financial Development in the European Union
by: Hyun‐Jung Nam, et al.
Published: (2025)
by: Hyun‐Jung Nam, et al.
Published: (2025)
FLOWER: Flow-Based Estimated Gaussian Guidance for General Speech Restoration
by: Yang, Da-Hee, et al.
Published: (2025)
by: Yang, Da-Hee, et al.
Published: (2025)
CIPHER: Counterfeit Image Pattern High-level Examination via Representation
by: Kim, Kyeonghun, et al.
Published: (2026)
by: Kim, Kyeonghun, et al.
Published: (2026)
DEX-TTS: Diffusion-based EXpressive Text-to-Speech with Style Modeling on Time Variability
by: Park, Hyun Joon, et al.
Published: (2024)
by: Park, Hyun Joon, et al.
Published: (2024)
Single‐Switch Forward‐Flyback High Step‐Up DC‐DC Converter With Reduced Switching Losses
by: Jae‐Won Yang, et al.
Published: (2026)
by: Jae‐Won Yang, et al.
Published: (2026)
Empowering Dysarthric Speech: Leveraging Advanced LLMs for Accurate Speech Correction and Multimodal Emotion Analysis
by: Attaluri, Kaushal, et al.
Published: (2024)
by: Attaluri, Kaushal, et al.
Published: (2024)
Seeing Speech and Sound: Distinguishing and Locating Audios in Visual Scenes
by: Ryu, Hyeonggon, et al.
Published: (2025)
by: Ryu, Hyeonggon, et al.
Published: (2025)
KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs
by: Kim, Haechan, et al.
Published: (2026)
by: Kim, Haechan, et al.
Published: (2026)
Improving Noise Robust Audio-Visual Speech Recognition via Router-Gated Cross-Modal Feature Fusion
by: Lim, DongHoon, et al.
Published: (2025)
by: Lim, DongHoon, et al.
Published: (2025)
Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics
by: Chae-Yeon, Lee, et al.
Published: (2025)
by: Chae-Yeon, Lee, et al.
Published: (2025)
A Dual-Branch Parallel Network for Speech Enhancement and Restoration
by: Yang, Da-Hee, et al.
Published: (2024)
by: Yang, Da-Hee, et al.
Published: (2024)
CLASS-IT: Conversational and Lecture-Aligned Small-Scale Instruction Tuning for BabyLMs
by: Capone, Luca, et al.
Published: (2025)
by: Capone, Luca, et al.
Published: (2025)
Macroscopic Monograin Nano‐Gyroid Thin Films in 4‐inch Scale Through Shear‐Rolling and Subsequent Solvent Annealing
by: Woo Hyun Nam, et al.
Published: (2024)
by: Woo Hyun Nam, et al.
Published: (2024)
Clustering and Mining Accented Speech for Inclusive and Fair Speech Recognition
by: Kim, Jaeyoung, et al.
Published: (2024)
by: Kim, Jaeyoung, et al.
Published: (2024)
Evaluation serum soluble interleukin 2 receptor with diagnosis and prognosis in canine solid tumour: 34 cases
by: Hyun NamKung, et al.
Published: (2024)
by: Hyun NamKung, et al.
Published: (2024)
End-to-end Automatic Speech Recognition and Speech Translation: Integration of Speech Foundational Models and LLMs
by: Luu, Nam, et al.
Published: (2025)
by: Luu, Nam, et al.
Published: (2025)
VietSuperSpeech: A Large-Scale Vietnamese Conversational Speech Dataset for ASR Fine-Tuning in Chatbot, Customer Support, and Call Center Applications
by: Do, Loan, et al.
Published: (2026)
by: Do, Loan, et al.
Published: (2026)
RapFlow-TTS: Rapid and High-Fidelity Text-to-Speech with Improved Consistency Flow Matching
by: Park, Hyun Joon, et al.
Published: (2025)
by: Park, Hyun Joon, et al.
Published: (2025)
Late Fusion and Multi-Level Fission Amplify Cross-Modal Transfer in Text-Speech LMs
by: Cuervo, Santiago, et al.
Published: (2025)
by: Cuervo, Santiago, et al.
Published: (2025)
Strategic Buried Defect Passivation of Perovskite Emitting Layers by Guanidinium Chloride for High‐Performance Pure Blue Perovskite Light Emitting Diodes
by: Jung Jae Do, et al.
Published: (2024)
by: Jung Jae Do, et al.
Published: (2024)
Dub-S2ST: Textless Speech-to-Speech Translation for Seamless Dubbing
by: Choi, Jeongsoo, et al.
Published: (2025)
by: Choi, Jeongsoo, et al.
Published: (2025)
Polyfactoriality as a Defining Criterion of Formulaic Speech
by: Schmale, Günter
Published: (2010)
by: Schmale, Günter
Published: (2010)
Translating the Optimized Durability of Co‐Based Anode Catalyst into Sustainable Anion Exchange Membrane Water Electrolysis
by: Sanghwi Han, et al.
Published: (2024)
by: Sanghwi Han, et al.
Published: (2024)
Vertification of Clinical Effectiveness of Dual‐task Cognitive Training Program and Clinical tDCS system
by: Jong‐Hyeon Kim, et al.
Published: (2025)
by: Jong‐Hyeon Kim, et al.
Published: (2025)
Application of co‐feeds in the catalytic conversion of light alkanes to aromatics: A comprehensive review
by: Yong Hyun Lim, et al.
Published: (2024)
by: Yong Hyun Lim, et al.
Published: (2024)
Similar Items
-
MathReader : Text-to-Speech for Mathematical Documents
by: Hyeon, Sieun, et al.
Published: (2025) -
MathBridge: A Large Corpus Dataset for Translating Spoken Mathematical Expressions into $LaTeX$ Formulas for Improved Readability
by: Jung, Kyudan, et al.
Published: (2024) -
Enhancing ASR Performance through OCR Word Frequency Analysis: Theoretical Foundations
by: Jung, Kyudan, et al.
Published: (2024) -
Is Retraining-Free Enough? The Necessity of Router Calibration for Efficient MoE Compression
by: Hyeon, Sieun, et al.
Published: (2026) -
TeXBLEU: Automatic Metric for Evaluate LaTeX Format
by: Jung, Kyudan, et al.
Published: (2024)