DoCIA: An Online Document-Level Context Incorporation Agent for Speech Translation
Fuente:
arXiv
Saved in:
| Main Authors: | Lyu, Xinglin, Tang, Wei, Li, Yuang, Zhao, Xiaofeng, Zhu, Ming, Li, Junhui, Lu, Yunfei, Zhang, Min, Wei, Daimeng, Yang, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-Preference Learning for Sentence-Level and Context-Aware Machine Translation
by: Li, Ying, et al.
Published: (2026)
by: Li, Ying, et al.
Published: (2026)
Two Intermediate Translations Are Better Than One: Fine-tuning LLMs for Document-level Translation Refinement
by: Dong, Yichen, et al.
Published: (2025)
by: Dong, Yichen, et al.
Published: (2025)
Improving LLM-based Document-level Machine Translation with Multi-Knowledge Fusion
by: Liu, Bin, et al.
Published: (2025)
by: Liu, Bin, et al.
Published: (2025)
DeMPT: Decoding-enhanced Multi-phase Prompt Tuning for Making LLMs Be Better Context-aware Translators
by: Lyu, Xinglin, et al.
Published: (2024)
by: Lyu, Xinglin, et al.
Published: (2024)
Loong: A Human-Like Long Document Translation Agent with Observe-and-Act Adaptive Context Selection
by: Wang, Yutong, et al.
Published: (2026)
by: Wang, Yutong, et al.
Published: (2026)
Cross-Domain Audio Deepfake Detection: Dataset and Analysis
by: Li, Yuang, et al.
Published: (2024)
by: Li, Yuang, et al.
Published: (2024)
Speech Translation Refinement using Large Language Models
by: Dou, Huaixia, et al.
Published: (2025)
by: Dou, Huaixia, et al.
Published: (2025)
Enhancing Document-level Translation of Large Language Model via Translation Mixed-instructions
by: Li, Yachao, et al.
Published: (2024)
by: Li, Yachao, et al.
Published: (2024)
Investigating Numerical Translation with Large Language Models
by: Tang, Wei, et al.
Published: (2025)
by: Tang, Wei, et al.
Published: (2025)
DelTA: An Online Document-Level Translation Agent Based on Multi-Level Memory
by: Wang, Yutong, et al.
Published: (2024)
by: Wang, Yutong, et al.
Published: (2024)
LLM with Relation Classifier for Document-Level Relation Extraction
by: Li, Xingzuo, et al.
Published: (2024)
by: Li, Xingzuo, et al.
Published: (2024)
Context-aware and Style-related Incremental Decoding framework for Discourse-Level Literary Translation
by: Luo, Yuanchang, et al.
Published: (2024)
by: Luo, Yuanchang, et al.
Published: (2024)
Towards Fine-Grained Code-Switch Speech Translation with Semantic Space Alignment
by: Gao, Yan, et al.
Published: (2025)
by: Gao, Yan, et al.
Published: (2025)
Large Language Model Should Understand Pinyin for Chinese ASR Error Correction
by: Li, Yuang, et al.
Published: (2024)
by: Li, Yuang, et al.
Published: (2024)
Locate-and-Focus: Enhancing Terminology Translation in Speech Language Models
by: Wu, Suhang, et al.
Published: (2025)
by: Wu, Suhang, et al.
Published: (2025)
Enhancing Speech Large Language Models with Prompt-Aware Mixture of Audio Encoders
by: Shan, Weiqiao, et al.
Published: (2025)
by: Shan, Weiqiao, et al.
Published: (2025)
A Multitask Training Approach to Enhance Whisper with Contextual Biasing and Open-Vocabulary Keyword Spotting
by: Li, Yuang, et al.
Published: (2023)
by: Li, Yuang, et al.
Published: (2023)
Align-then-Slide: A complete evaluation framework for Ultra-Long Document-Level Machine Translation
by: Guo, Jiaxin, et al.
Published: (2025)
by: Guo, Jiaxin, et al.
Published: (2025)
Automatic Evaluation Metrics for Document-level Translation: Overview, Challenges and Trends
by: GUO, Jiaxin, et al.
Published: (2025)
by: GUO, Jiaxin, et al.
Published: (2025)
Doc-Guided Sent2Sent++: A Sent2Sent++ Agent with Doc-Guided memory for Document-level Machine Translation
by: Guo, Jiaxin, et al.
Published: (2025)
by: Guo, Jiaxin, et al.
Published: (2025)
Optimizing Speech Multi-View Feature Fusion through Conditional Computation
by: Shan, Weiqiao, et al.
Published: (2025)
by: Shan, Weiqiao, et al.
Published: (2025)
R-BI: Regularized Batched Inputs enhance Incremental Decoding Framework for Low-Latency Simultaneous Speech Translation
by: Guo, Jiaxin, et al.
Published: (2024)
by: Guo, Jiaxin, et al.
Published: (2024)
A Novel Paradigm Boosting Translation Capabilities of Large Language Models
by: Guo, Jiaxin, et al.
Published: (2024)
by: Guo, Jiaxin, et al.
Published: (2024)
Hard-Synth: Synthesizing Diverse Hard Samples for ASR using Zero-Shot TTS and LLM
by: Yu, Jiawei, et al.
Published: (2024)
by: Yu, Jiawei, et al.
Published: (2024)
Analysis of Volatile Organic Compounds of Different Types of Peppers (Capsicum annuum L.) Using Comprehensive Two‐Dimensional Gas Chromatography With Time‐of‐Flight Mass Spectrometry
by: Huixia Zhu, et al.
Published: (2024)
by: Huixia Zhu, et al.
Published: (2024)
UCorrect: An Unsupervised Framework for Automatic Speech Recognition Error Correction
by: Guo, Jiaxin, et al.
Published: (2024)
by: Guo, Jiaxin, et al.
Published: (2024)
Unlocking Fine-Grained Translation Quality Estimation in LRMs through Synergistically Evolving Implicit and Explicit Reasoning
by: Dang, Renfei, et al.
Published: (2026)
by: Dang, Renfei, et al.
Published: (2026)
LA-RAG:Enhancing LLM-based ASR Accuracy with Retrieval-Augmented Generation
by: Li, Shaojun, et al.
Published: (2024)
by: Li, Shaojun, et al.
Published: (2024)
Why Not Transform Chat Large Language Models to Non-English?
by: Geng, Xiang, et al.
Published: (2024)
by: Geng, Xiang, et al.
Published: (2024)
Function-to-Style Guidance of LLMs for Code Translation
by: Zhang, Longhui, et al.
Published: (2025)
by: Zhang, Longhui, et al.
Published: (2025)
L-CiteEval: Do Long-Context Models Truly Leverage Context for Responding?
by: Tang, Zecheng, et al.
Published: (2024)
by: Tang, Zecheng, et al.
Published: (2024)
An End-to-End Speech Summarization Using Large Language Model
by: Shang, Hengchao, et al.
Published: (2024)
by: Shang, Hengchao, et al.
Published: (2024)
Exploring In-Context Learning of Textless Speech Language Model for Speech Classification Tasks
by: Hsu, Ming-Hao, et al.
Published: (2023)
by: Hsu, Ming-Hao, et al.
Published: (2023)
GLIMMER: Incorporating Graph and Lexical Features in Unsupervised Multi-Document Summarization
by: Liu, Ran, et al.
Published: (2024)
by: Liu, Ran, et al.
Published: (2024)
StreamSpeech: Simultaneous Speech-to-Speech Translation with Multi-task Learning
by: Zhang, Shaolei, et al.
Published: (2024)
by: Zhang, Shaolei, et al.
Published: (2024)
HW-TSC's Submission to the CCMT 2024 Machine Translation Tasks
by: Wu, Zhanglin, et al.
Published: (2024)
by: Wu, Zhanglin, et al.
Published: (2024)
SpeechEE: A Novel Benchmark for Speech Event Extraction
by: Wang, Bin, et al.
Published: (2024)
by: Wang, Bin, et al.
Published: (2024)
Can We Achieve High-quality Direct Speech-to-Speech Translation without Parallel Speech Data?
by: Fang, Qingkai, et al.
Published: (2024)
by: Fang, Qingkai, et al.
Published: (2024)
CTC-based Non-autoregressive Textless Speech-to-Speech Translation
by: Fang, Qingkai, et al.
Published: (2024)
by: Fang, Qingkai, et al.
Published: (2024)
A Study on Incorporating Whisper for Robust Speech Assessment
by: Zezario, Ryandhimas E., et al.
Published: (2023)
by: Zezario, Ryandhimas E., et al.
Published: (2023)
Similar Items
-
Cross-Preference Learning for Sentence-Level and Context-Aware Machine Translation
by: Li, Ying, et al.
Published: (2026) -
Two Intermediate Translations Are Better Than One: Fine-tuning LLMs for Document-level Translation Refinement
by: Dong, Yichen, et al.
Published: (2025) -
Improving LLM-based Document-level Machine Translation with Multi-Knowledge Fusion
by: Liu, Bin, et al.
Published: (2025) -
DeMPT: Decoding-enhanced Multi-phase Prompt Tuning for Making LLMs Be Better Context-aware Translators
by: Lyu, Xinglin, et al.
Published: (2024) -
Loong: A Human-Like Long Document Translation Agent with Observe-and-Act Adaptive Context Selection
by: Wang, Yutong, et al.
Published: (2026)