SimulSense: Sense-Driven Interpreting for Efficient Simultaneous Speech Translation
Fuente:
arXiv
Saved in:
| Main Authors: | Tan, Haotian, Ouchi, Hiroki, Sakti, Sakriani |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Contrastive Feedback Mechanism for Simultaneous Speech Translation
by: Tan, Haotian, et al.
Published: (2024)
by: Tan, Haotian, et al.
Published: (2024)
SC-SOT: Conditioning the Decoder on Diarized Speaker Information for End-to-End Overlapped Speech Recognition
by: Hirano, Yuta, et al.
Published: (2025)
by: Hirano, Yuta, et al.
Published: (2025)
NAIST Simultaneous Speech Translation System for IWSLT 2024
by: Ko, Yuka, et al.
Published: (2024)
by: Ko, Yuka, et al.
Published: (2024)
On the Problem of Text-To-Speech Model Selection for Synthetic Data Generation in Automatic Speech Recognition
by: Rossenbach, Nick, et al.
Published: (2024)
by: Rossenbach, Nick, et al.
Published: (2024)
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
by: Adila, Aulia, et al.
Published: (2024)
by: Adila, Aulia, et al.
Published: (2024)
Indonesian-English Code-Switching Speech Synthesizer Utilizing Multilingual STEN-TTS and Bert LID
by: Handoyo, Ahmad Alfani, et al.
Published: (2024)
by: Handoyo, Ahmad Alfani, et al.
Published: (2024)
Continual Learning in Machine Speech Chain Using Gradient Episodic Memory
by: Tyndall, Geoffrey, et al.
Published: (2024)
by: Tyndall, Geoffrey, et al.
Published: (2024)
SimulTron: On-Device Simultaneous Speech to Speech Translation
by: Agranovich, Alex, et al.
Published: (2024)
by: Agranovich, Alex, et al.
Published: (2024)
ATD-Trans: A Geographically Grounded Japanese-English Travelogue Translation Dataset
by: Higashiyama, Shohei, et al.
Published: (2026)
by: Higashiyama, Shohei, et al.
Published: (2026)
Efficient and Adaptive Simultaneous Speech Translation with Fully Unidirectional Architecture
by: Fu, Biao, et al.
Published: (2025)
by: Fu, Biao, et al.
Published: (2025)
SimulS2S-LLM: Unlocking Simultaneous Inference of Speech LLMs for Speech-to-Speech Translation
by: Deng, Keqi, et al.
Published: (2025)
by: Deng, Keqi, et al.
Published: (2025)
Conversational SimulMT: Efficient Simultaneous Translation with Large Language Models
by: Wang, Minghan, et al.
Published: (2024)
by: Wang, Minghan, et al.
Published: (2024)
SimulSeamless: FBK at IWSLT 2024 Simultaneous Speech Translation
by: Papi, Sara, et al.
Published: (2024)
by: Papi, Sara, et al.
Published: (2024)
High-Fidelity Simultaneous Speech-To-Speech Translation
by: Labiausse, Tom, et al.
Published: (2025)
by: Labiausse, Tom, et al.
Published: (2025)
CMU's IWSLT 2025 Simultaneous Speech Translation System
by: Ouyang, Siqi, et al.
Published: (2025)
by: Ouyang, Siqi, et al.
Published: (2025)
Hierarchical Policy Optimization for Simultaneous Translation of Unbounded Speech
by: Ouyang, Siqi, et al.
Published: (2026)
by: Ouyang, Siqi, et al.
Published: (2026)
SimulU: Training-free Policy for Long-form Simultaneous Speech-to-Speech Translation
by: Djanibekov, Amirbek, et al.
Published: (2026)
by: Djanibekov, Amirbek, et al.
Published: (2026)
Toward Natural Emotional Text-To-Speech System with Fine-Grained Non-Verbal Expression Control
by: Zhou, Wangzixi, et al.
Published: (2026)
by: Zhou, Wangzixi, et al.
Published: (2026)
Simultaneous Speech-to-Speech Translation Without Aligned Data
by: Labiausse, Tom, et al.
Published: (2026)
by: Labiausse, Tom, et al.
Published: (2026)
Continuous Rating as Reliable Human Evaluation of Simultaneous Speech Translation
by: Javorský, Dávid, et al.
Published: (2022)
by: Javorský, Dávid, et al.
Published: (2022)
MLLP-VRAIN UPV system for the IWSLT 2025 Simultaneous Speech Translation Translation task
by: Iranzo-Sánchez, Jorge, et al.
Published: (2025)
by: Iranzo-Sánchez, Jorge, et al.
Published: (2025)
FASST: Fast LLM-based Simultaneous Speech Translation
by: Ouyang, Siqi, et al.
Published: (2024)
by: Ouyang, Siqi, et al.
Published: (2024)
CMU's IWSLT 2024 Simultaneous Speech Translation System
by: Xu, Xi, et al.
Published: (2024)
by: Xu, Xi, et al.
Published: (2024)
SimulMEGA: MoE Routers are Advanced Policy Makers for Simultaneous Speech Translation
by: Le, Chenyang, et al.
Published: (2025)
by: Le, Chenyang, et al.
Published: (2025)
DPO-Tuned Large Language Models for Segmentation in Simultaneous Speech Translation
by: Yang, Zeyu, et al.
Published: (2025)
by: Yang, Zeyu, et al.
Published: (2025)
SASST: Leveraging Syntax-Aware Chunking and LLMs for Simultaneous Speech Translation
by: Yang, Zeyu, et al.
Published: (2025)
by: Yang, Zeyu, et al.
Published: (2025)
RASST: Fast Cross-modal Retrieval-Augmented Simultaneous Speech Translation
by: Luo, Jiaxuan, et al.
Published: (2026)
by: Luo, Jiaxuan, et al.
Published: (2026)
Exploring the Correlation between Human and Machine Evaluation of Simultaneous Speech Translation
by: Wang, Xiaoman, et al.
Published: (2024)
by: Wang, Xiaoman, et al.
Published: (2024)
Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice
by: Cheng, Shanbo, et al.
Published: (2025)
by: Cheng, Shanbo, et al.
Published: (2025)
REINA: Regularized Entropy Information-Based Loss for Efficient Simultaneous Speech Translation
by: Hirschkind, Nameer, et al.
Published: (2025)
by: Hirschkind, Nameer, et al.
Published: (2025)
NERsocial: Efficient Named Entity Recognition Dataset Construction for Human-Robot Interaction Utilizing RapidNER
by: Atuhurra, Jesse, et al.
Published: (2024)
by: Atuhurra, Jesse, et al.
Published: (2024)
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
by: Huber, Christian, et al.
Published: (2023)
by: Huber, Christian, et al.
Published: (2023)
Simultaneous Translation with Offline Speech and LLM Models in CUNI Submission to IWSLT 2025
by: Macháček, Dominik, et al.
Published: (2025)
by: Macháček, Dominik, et al.
Published: (2025)
Redefining Machine Simultaneous Interpretation: From Incremental Translation to Human-Like Strategies
by: Zhang, Qianen, et al.
Published: (2025)
by: Zhang, Qianen, et al.
Published: (2025)
Redefining Machine Simultaneous Interpretation: From Incremental Translation to Human-Like Strategies
by: Zhang, Qianen, et al.
Published: (2026)
by: Zhang, Qianen, et al.
Published: (2026)
StreamSpeech: Simultaneous Speech-to-Speech Translation with Multi-task Learning
by: Zhang, Shaolei, et al.
Published: (2024)
by: Zhang, Shaolei, et al.
Published: (2024)
MobQA: A Benchmark Dataset for Semantic Understanding of Human Mobility Data through Question Answering
by: Asano, Hikaru, et al.
Published: (2025)
by: Asano, Hikaru, et al.
Published: (2025)
Text2Traj2Text: Learning-by-Synthesis Framework for Contextual Captioning of Human Movement Trajectories
by: Asano, Hikaru, et al.
Published: (2024)
by: Asano, Hikaru, et al.
Published: (2024)
iBERT: Interpretable Embeddings via Sense Decomposition
by: Anand, Vishal, et al.
Published: (2025)
by: Anand, Vishal, et al.
Published: (2025)
InfiniSST: Simultaneous Translation of Unbounded Speech with Large Language Model
by: Ouyang, Siqi, et al.
Published: (2025)
by: Ouyang, Siqi, et al.
Published: (2025)
Similar Items
-
Contrastive Feedback Mechanism for Simultaneous Speech Translation
by: Tan, Haotian, et al.
Published: (2024) -
SC-SOT: Conditioning the Decoder on Diarized Speaker Information for End-to-End Overlapped Speech Recognition
by: Hirano, Yuta, et al.
Published: (2025) -
NAIST Simultaneous Speech Translation System for IWSLT 2024
by: Ko, Yuka, et al.
Published: (2024) -
On the Problem of Text-To-Speech Model Selection for Synthetic Data Generation in Automatic Speech Recognition
by: Rossenbach, Nick, et al.
Published: (2024) -
Enhancing Indonesian Automatic Speech Recognition: Evaluating Multilingual Models with Diverse Speech Variabilities
by: Adila, Aulia, et al.
Published: (2024)