Testing LLMs' Capabilities in Annotating Translations Based on an Error Typology Designed for LSP Translation: First Experiments with ChatGPT
Fuente:
arXiv
Guardado en:
| Autores principales: | Minder, Joachim, Wisniewski, Guillaume, Kübler, Natalie |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection
por: Amiri, Mahdi, et al.
Publicado: (2025)
por: Amiri, Mahdi, et al.
Publicado: (2025)
Zero-resource Speech Translation and Recognition with LLMs
por: Mundnich, Karel, et al.
Publicado: (2024)
por: Mundnich, Karel, et al.
Publicado: (2024)
LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT
por: Zhuo, Le, et al.
Publicado: (2023)
por: Zhuo, Le, et al.
Publicado: (2023)
Transcribing and Translating, Fast and Slow: Joint Speech Translation and Recognition
por: Moritz, Niko, et al.
Publicado: (2024)
por: Moritz, Niko, et al.
Publicado: (2024)
Scheduled Interleaved Speech-Text Training for Speech-to-Speech Translation with LLMs
por: Futami, Hayato, et al.
Publicado: (2025)
por: Futami, Hayato, et al.
Publicado: (2025)
Establishing degrees of closeness between audio recordings along different dimensions using large-scale cross-lingual models
por: Fily, Maxime, et al.
Publicado: (2024)
por: Fily, Maxime, et al.
Publicado: (2024)
Spatial Speech Translation: Translating Across Space With Binaural Hearables
por: Chen, Tuochao, et al.
Publicado: (2025)
por: Chen, Tuochao, et al.
Publicado: (2025)
WavCaps: A ChatGPT-Assisted Weakly-Labelled Audio Captioning Dataset for Audio-Language Multimodal Research
por: Mei, Xinhao, et al.
Publicado: (2023)
por: Mei, Xinhao, et al.
Publicado: (2023)
SimulS2S-LLM: Unlocking Simultaneous Inference of Speech LLMs for Speech-to-Speech Translation
por: Deng, Keqi, et al.
Publicado: (2025)
por: Deng, Keqi, et al.
Publicado: (2025)
Textless Speech-to-Speech Translation With Limited Parallel Data
por: Diwan, Anuj, et al.
Publicado: (2023)
por: Diwan, Anuj, et al.
Publicado: (2023)
Translating speech with just images
por: Oneata, Dan, et al.
Publicado: (2024)
por: Oneata, Dan, et al.
Publicado: (2024)
Granary: Speech Recognition and Translation Dataset in 25 European Languages
por: Koluguri, Nithin Rao, et al.
Publicado: (2025)
por: Koluguri, Nithin Rao, et al.
Publicado: (2025)
Assessing the Impact of Anisotropy in Neural Representations of Speech: A Case Study on Keyword Spotting
por: Wisniewski, Guillaume, et al.
Publicado: (2025)
por: Wisniewski, Guillaume, et al.
Publicado: (2025)
Finetuning End-to-End Models for Estonian Conversational Spoken Language Translation
por: Sildam, Tiia, et al.
Publicado: (2024)
por: Sildam, Tiia, et al.
Publicado: (2024)
Label-Synchronous Neural Transducer for E2E Simultaneous Speech Translation
por: Deng, Keqi, et al.
Publicado: (2024)
por: Deng, Keqi, et al.
Publicado: (2024)
High-Fidelity Simultaneous Speech-To-Speech Translation
por: Labiausse, Tom, et al.
Publicado: (2025)
por: Labiausse, Tom, et al.
Publicado: (2025)
Representation Purification for End-to-End Speech Translation
por: Zhang, Chengwei, et al.
Publicado: (2024)
por: Zhang, Chengwei, et al.
Publicado: (2024)
Direct Speech to Speech Translation: A Review
por: Sarim, Mohammad, et al.
Publicado: (2025)
por: Sarim, Mohammad, et al.
Publicado: (2025)
A Multi-Dialectal Dataset for German Dialect ASR and Dialect-to-Standard Speech Translation
por: Blaschke, Verena, et al.
Publicado: (2025)
por: Blaschke, Verena, et al.
Publicado: (2025)
Longer is (Not Necessarily) Stronger: Punctuated Long-Sequence Training for Enhanced Speech Recognition and Translation
por: Koluguri, Nithin Rao, et al.
Publicado: (2024)
por: Koluguri, Nithin Rao, et al.
Publicado: (2024)
End-to-End Speech Translation for Low-Resource Languages Using Weakly Labeled Data
por: Pothula, Aishwarya, et al.
Publicado: (2025)
por: Pothula, Aishwarya, et al.
Publicado: (2025)
TeluguST-46: A Benchmark Corpus and Comprehensive Evaluation for Telugu-English Speech Translation
por: Akkiraju, Bhavana, et al.
Publicado: (2025)
por: Akkiraju, Bhavana, et al.
Publicado: (2025)
HENT-SRT: Hierarchical Efficient Neural Transducer with Self-Distillation for Joint Speech Recognition and Translation
por: Hussein, Amir, et al.
Publicado: (2025)
por: Hussein, Amir, et al.
Publicado: (2025)
Simultaneous Speech-to-Speech Translation Without Aligned Data
por: Labiausse, Tom, et al.
Publicado: (2026)
por: Labiausse, Tom, et al.
Publicado: (2026)
Lightweight Audio Segmentation for Long-form Speech Translation
por: Lee, Jaesong, et al.
Publicado: (2024)
por: Lee, Jaesong, et al.
Publicado: (2024)
EMMeTT: Efficient Multimodal Machine Translation Training
por: Żelasko, Piotr, et al.
Publicado: (2024)
por: Żelasko, Piotr, et al.
Publicado: (2024)
End-to-End Speech-to-Text Translation: A Survey
por: Sethiya, Nivedita, et al.
Publicado: (2023)
por: Sethiya, Nivedita, et al.
Publicado: (2023)
DiariST: Streaming Speech Translation with Speaker Diarization
por: Yang, Mu, et al.
Publicado: (2023)
por: Yang, Mu, et al.
Publicado: (2023)
NAIST Simultaneous Speech Translation System for IWSLT 2024
por: Ko, Yuka, et al.
Publicado: (2024)
por: Ko, Yuka, et al.
Publicado: (2024)
Evolutionary Prompt Design for LLM-Based Post-ASR Error Correction
por: Sachdev, Rithik, et al.
Publicado: (2024)
por: Sachdev, Rithik, et al.
Publicado: (2024)
Direct Speech-to-Speech Neural Machine Translation: A Survey
por: Gupta, Mahendra, et al.
Publicado: (2024)
por: Gupta, Mahendra, et al.
Publicado: (2024)
Efficient Speech Translation through Model Compression and Knowledge Distillation
por: Moslem, Yasmin
Publicado: (2025)
por: Moslem, Yasmin
Publicado: (2025)
Word Level Timestamp Generation for Automatic Speech Recognition and Translation
por: Hu, Ke, et al.
Publicado: (2025)
por: Hu, Ke, et al.
Publicado: (2025)
Attempt Towards Stress Transfer in Speech-to-Speech Machine Translation
por: Akarsh, Sai, et al.
Publicado: (2024)
por: Akarsh, Sai, et al.
Publicado: (2024)
Direct Simultaneous Translation Activation for Large Audio-Language Models
por: Zhang, Pei, et al.
Publicado: (2025)
por: Zhang, Pei, et al.
Publicado: (2025)
Isochrony-Controlled Speech-to-Text Translation: A study on translating from Sino-Tibetan to Indo-European Languages
por: Yousefi, Midia, et al.
Publicado: (2024)
por: Yousefi, Midia, et al.
Publicado: (2024)
Smooth Operators: LLMs Translating Imperfect Hints into Disfluency-Rich Transcripts
por: Altinok, Duygu
Publicado: (2025)
por: Altinok, Duygu
Publicado: (2025)
Textless Unit-to-Unit training for Many-to-Many Multilingual Speech-to-Speech Translation
por: Kim, Minsu, et al.
Publicado: (2023)
por: Kim, Minsu, et al.
Publicado: (2023)
Dub-S2ST: Textless Speech-to-Speech Translation for Seamless Dubbing
por: Choi, Jeongsoo, et al.
Publicado: (2025)
por: Choi, Jeongsoo, et al.
Publicado: (2025)
Compact Speech Translation Models via Discrete Speech Units Pretraining
por: Lam, Tsz Kin, et al.
Publicado: (2024)
por: Lam, Tsz Kin, et al.
Publicado: (2024)
Ejemplares similares
-
Exploring In-Context Learning Capabilities of ChatGPT for Pathological Speech Detection
por: Amiri, Mahdi, et al.
Publicado: (2025) -
Zero-resource Speech Translation and Recognition with LLMs
por: Mundnich, Karel, et al.
Publicado: (2024) -
LyricWhiz: Robust Multilingual Zero-shot Lyrics Transcription by Whispering to ChatGPT
por: Zhuo, Le, et al.
Publicado: (2023) -
Transcribing and Translating, Fast and Slow: Joint Speech Translation and Recognition
por: Moritz, Niko, et al.
Publicado: (2024) -
Scheduled Interleaved Speech-Text Training for Speech-to-Speech Translation with LLMs
por: Futami, Hayato, et al.
Publicado: (2025)