FormalASR: End-to-End Spoken Chinese to Formal Text
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ning, Wanyi, Guo, Yinshang, Qian, Haitao, Cheng, Jiyuan, Feng, Weiyuan, Zhang, Yufei |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Using Large Language Model for End-to-End Chinese ASR and NER
par: Li, Yuang, et autres
Publié: (2024)
par: Li, Yuang, et autres
Publié: (2024)
GSQA: An End-to-End Model for Generative Spoken Question Answering
par: Shih, Min-Han, et autres
Publié: (2023)
par: Shih, Min-Han, et autres
Publié: (2023)
Retrieval Augmented End-to-End Spoken Dialog Models
par: Wang, Mingqiu, et autres
Publié: (2024)
par: Wang, Mingqiu, et autres
Publié: (2024)
The End of Manual Decoding: Towards Truly End-to-End Language Models
par: Wang, Zhichao, et autres
Publié: (2025)
par: Wang, Zhichao, et autres
Publié: (2025)
FormalML: A Benchmark for Evaluating Formal Subgoal Completion in Machine Learning Theory
par: Yang, Xiao-Wen, et autres
Publié: (2025)
par: Yang, Xiao-Wen, et autres
Publié: (2025)
From Informal to Formal -- Incorporating and Evaluating LLMs on Natural Language Requirements to Verifiable Formal Proofs
par: Cao, Jialun, et autres
Publié: (2025)
par: Cao, Jialun, et autres
Publié: (2025)
Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models
par: Hsiao, Chi-Yuan, et autres
Publié: (2025)
par: Hsiao, Chi-Yuan, et autres
Publié: (2025)
Let's Reason Formally: Natural-Formal Hybrid Reasoning Enhances LLM's Math Capability
par: Wang, Ruida, et autres
Publié: (2025)
par: Wang, Ruida, et autres
Publié: (2025)
MOSS-TTSD: Text to Spoken Dialogue Generation
par: Zhang, Yuqian, et autres
Publié: (2026)
par: Zhang, Yuqian, et autres
Publié: (2026)
Generalizable End-to-End Tool-Use RL with Synthetic CodeGym
par: Du, Weihua, et autres
Publié: (2025)
par: Du, Weihua, et autres
Publié: (2025)
End-to-End Graph Flattening Method for Large Language Models
par: Hong, Bin, et autres
Publié: (2024)
par: Hong, Bin, et autres
Publié: (2024)
The Speech-LLM Takes It All: A Truly Fully End-to-End Spoken Dialogue State Tracking Approach
par: Ghazal, Nizar El, et autres
Publié: (2025)
par: Ghazal, Nizar El, et autres
Publié: (2025)
Zero-Shot End-to-End Relation Extraction in Chinese: A Comparative Study of Gemini, LLaMA and ChatGPT
par: Du, Shaoshuai, et autres
Publié: (2025)
par: Du, Shaoshuai, et autres
Publié: (2025)
Distilling an End-to-End Voice Assistant Without Instruction Training Data
par: Held, William, et autres
Publié: (2024)
par: Held, William, et autres
Publié: (2024)
Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation
par: Lan, Zhibin, et autres
Publié: (2024)
par: Lan, Zhibin, et autres
Publié: (2024)
MedSpeak: A Knowledge Graph-Aided ASR Error Correction Framework for Spoken Medical QA
par: Song, Yutong, et autres
Publié: (2026)
par: Song, Yutong, et autres
Publié: (2026)
Language Model Inversion through End-to-End Differentiation
par: Denamganaï, Kevin Yandoka, et autres
Publié: (2026)
par: Denamganaï, Kevin Yandoka, et autres
Publié: (2026)
End-to-End Aspect-Guided Review Summarization at Scale
par: Boytsov, Ilya, et autres
Publié: (2025)
par: Boytsov, Ilya, et autres
Publié: (2025)
A Non-autoregressive Generation Framework for End-to-End Simultaneous Speech-to-Speech Translation
par: Ma, Zhengrui, et autres
Publié: (2024)
par: Ma, Zhengrui, et autres
Publié: (2024)
MedGo: A Chinese Medical Large Language Model
par: Zhang, Haitao, et autres
Publié: (2024)
par: Zhang, Haitao, et autres
Publié: (2024)
Formality Style Transfer in Persian
par: Falakaflaki, Parastoo, et autres
Publié: (2024)
par: Falakaflaki, Parastoo, et autres
Publié: (2024)
Formalizing Style in Personal Narratives
par: Cortal, Gustave, et autres
Publié: (2025)
par: Cortal, Gustave, et autres
Publié: (2025)
Dynamic Long Context Reasoning over Compressed Memory via End-to-End Reinforcement Learning
par: Chen, Zhuoen, et autres
Publié: (2026)
par: Chen, Zhuoen, et autres
Publié: (2026)
AD-AGENT: A Multi-agent Framework for End-to-end Anomaly Detection
par: Yang, Tiankai, et autres
Publié: (2025)
par: Yang, Tiankai, et autres
Publié: (2025)
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
par: He, Hongliang, et autres
Publié: (2024)
par: He, Hongliang, et autres
Publié: (2024)
Iterative Formalization and Planning in Partially Observable Environments
par: Gong, Liancheng, et autres
Publié: (2025)
par: Gong, Liancheng, et autres
Publié: (2025)
End-to-End Long Document Summarization using Gradient Caching
par: Saxena, Rohit, et autres
Publié: (2025)
par: Saxena, Rohit, et autres
Publié: (2025)
Meow: End-to-End Outline Writing for Automatic Academic Survey
par: Ma, Zhaoyu, et autres
Publié: (2025)
par: Ma, Zhaoyu, et autres
Publié: (2025)
End-To-End Clinical Trial Matching with Large Language Models
par: Ferber, Dyke, et autres
Publié: (2024)
par: Ferber, Dyke, et autres
Publié: (2024)
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
par: Huber, Christian, et autres
Publié: (2023)
par: Huber, Christian, et autres
Publié: (2023)
VERGE: Formal Refinement and Guidance Engine for Verifiable LLM Reasoning
par: Singh, Vikash, et autres
Publié: (2026)
par: Singh, Vikash, et autres
Publié: (2026)
OProver: A Unified Framework for Agentic Formal Theorem Proving
par: Ma, David, et autres
Publié: (2026)
par: Ma, David, et autres
Publié: (2026)
LiveLongBench: Tackling Long-Context Understanding for Spoken Texts from Live Streams
par: Wu, Yongxuan, et autres
Publié: (2025)
par: Wu, Yongxuan, et autres
Publié: (2025)
MiCU: End-to-End Smart Home Command Understanding with Large Language Model
par: Han, Haowei, et autres
Publié: (2026)
par: Han, Haowei, et autres
Publié: (2026)
CT2C-QA: Multimodal Question Answering over Chinese Text, Table and Chart
par: Zhao, Bowen, et autres
Publié: (2024)
par: Zhao, Bowen, et autres
Publié: (2024)
E2E-AFG: An End-to-End Model with Adaptive Filtering for Retrieval-Augmented Generation
par: Jiang, Yun, et autres
Publié: (2024)
par: Jiang, Yun, et autres
Publié: (2024)
End-to-End Argument Mining through Autoregressive Argumentative Structure Prediction
par: Das, Nilmadhab, et autres
Publié: (2025)
par: Das, Nilmadhab, et autres
Publié: (2025)
WebDS: An End-to-End Benchmark for Web-based Data Science
par: Hsu, Ethan, et autres
Publié: (2025)
par: Hsu, Ethan, et autres
Publié: (2025)
SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents
par: Si, Shuzheng, et autres
Publié: (2023)
par: Si, Shuzheng, et autres
Publié: (2023)
Comparing Data Augmentation Methods for End-to-End Task-Oriented Dialog Systems
par: Vlachos, Christos, et autres
Publié: (2024)
par: Vlachos, Christos, et autres
Publié: (2024)
Documents similaires
-
Using Large Language Model for End-to-End Chinese ASR and NER
par: Li, Yuang, et autres
Publié: (2024) -
GSQA: An End-to-End Model for Generative Spoken Question Answering
par: Shih, Min-Han, et autres
Publié: (2023) -
Retrieval Augmented End-to-End Spoken Dialog Models
par: Wang, Mingqiu, et autres
Publié: (2024) -
The End of Manual Decoding: Towards Truly End-to-End Language Models
par: Wang, Zhichao, et autres
Publié: (2025) -
FormalML: A Benchmark for Evaluating Formal Subgoal Completion in Machine Learning Theory
par: Yang, Xiao-Wen, et autres
Publié: (2025)