Different Speech Translation Models Encode and Translate Speaker Gender Differently
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fucci, Dennis, Gaido, Marco, Negri, Matteo, Bentivogli, Luisa, Martins, Andre, Attanasio, Giuseppe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation
von: Conti, Lina, et al.
Veröffentlicht: (2025)
von: Conti, Lina, et al.
Veröffentlicht: (2025)
Speech Translation with Speech Foundation Models and Large Language Models: What is There and What is Missing?
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
Cross-Attention is Half Explanation in Speech-to-Text Models
von: Papi, Sara, et al.
Veröffentlicht: (2025)
von: Papi, Sara, et al.
Veröffentlicht: (2025)
How do Hyenas deal with Human Speech? Speech Recognition and Translation with ConfHyena
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
SimulSeamless: FBK at IWSLT 2024 Simultaneous Speech Translation
von: Papi, Sara, et al.
Veröffentlicht: (2024)
von: Papi, Sara, et al.
Veröffentlicht: (2024)
A Prompt Response to the Demand for Automatic Gender-Neutral Translation
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2024)
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2024)
StreamAtt: Direct Streaming Speech-to-Text Translation with Attention-based Audio History Selection
von: Papi, Sara, et al.
Veröffentlicht: (2024)
von: Papi, Sara, et al.
Veröffentlicht: (2024)
The Unheard Alternative: Contrastive Explanations for Speech-to-Text Models
von: Conti, Lina, et al.
Veröffentlicht: (2025)
von: Conti, Lina, et al.
Veröffentlicht: (2025)
Simulstream: Open-Source Toolkit for Evaluation and Demonstration of Streaming Speech-to-Text Translation Systems
von: Gaido, Marco, et al.
Veröffentlicht: (2025)
von: Gaido, Marco, et al.
Veröffentlicht: (2025)
How to Evaluate Speech Translation with Source-Aware Neural MT Metrics
von: Cettolo, Mauro, et al.
Veröffentlicht: (2025)
von: Cettolo, Mauro, et al.
Veröffentlicht: (2025)
Echoes of Phonetics: Unveiling Relevant Acoustic Cues for ASR via Feature Attribution
von: Fucci, Dennis, et al.
Veröffentlicht: (2025)
von: Fucci, Dennis, et al.
Veröffentlicht: (2025)
SPES: Spectrogram Perturbation for Explainable Speech-to-Text Generation
von: Fucci, Dennis, et al.
Veröffentlicht: (2024)
von: Fucci, Dennis, et al.
Veröffentlicht: (2024)
Enhancing Gender-Inclusive Machine Translation with Neomorphemes and Large Language Models
von: Piergentili, Andrea, et al.
Veröffentlicht: (2024)
von: Piergentili, Andrea, et al.
Veröffentlicht: (2024)
An LLM-as-a-judge Approach for Scalable Gender-Neutral Translation Evaluation
von: Piergentili, Andrea, et al.
Veröffentlicht: (2025)
von: Piergentili, Andrea, et al.
Veröffentlicht: (2025)
Translation in the Hands of Many:Centering Lay Users in Machine Translation Interactions
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2025)
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2025)
SimulU: Training-free Policy for Long-form Simultaneous Speech-to-Speech Translation
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2026)
von: Djanibekov, Amirbek, et al.
Veröffentlicht: (2026)
Mind the Inclusivity Gap: Multilingual Gender-Neutral Translation Evaluation with mGeNTE
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2025)
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2025)
Twists, Humps, and Pebbles: Multilingual Speech Recognition Models Exhibit Gender Performance Gaps
von: Attanasio, Giuseppe, et al.
Veröffentlicht: (2024)
von: Attanasio, Giuseppe, et al.
Veröffentlicht: (2024)
What the Harm? Quantifying the Tangible Impact of Gender Bias in Machine Translation with a Human-centered Study
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2024)
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2024)
SBAAM! Eliminating Transcript Dependency in Automatic Subtitling
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
Watching the Watchers: Exposing Gender Disparities in Machine Translation Quality Estimation
von: Zaranis, Emmanouil, et al.
Veröffentlicht: (2024)
von: Zaranis, Emmanouil, et al.
Veröffentlicht: (2024)
Gender-Neutral Rewriting in Italian: Models, Approaches, and Trade-offs
von: Piergentili, Andrea, et al.
Veröffentlicht: (2025)
von: Piergentili, Andrea, et al.
Veröffentlicht: (2025)
The Warmup Dilemma: How Learning Rate Strategies Impact Speech-to-Text Model Convergence
von: Gaido, Marco, et al.
Veröffentlicht: (2025)
von: Gaido, Marco, et al.
Veröffentlicht: (2025)
FAMA: The First Large-Scale Open-Science Speech Foundation Model for English and Italian
von: Papi, Sara, et al.
Veröffentlicht: (2025)
von: Papi, Sara, et al.
Veröffentlicht: (2025)
MOSEL: 950,000 Hours of Speech Data for Open-Source Speech Foundation Model Training on EU Languages
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
von: Gaido, Marco, et al.
Veröffentlicht: (2024)
Building Bridges: A Dataset for Evaluating Gender-Fair Machine Translation into German
von: Lardelli, Manuel, et al.
Veröffentlicht: (2024)
von: Lardelli, Manuel, et al.
Veröffentlicht: (2024)
DOA: Training-Free Decoder-Only Attention Policy for Long-Form Simultaneous Translation with SpeechLLMs
von: Papi, Sara, et al.
Veröffentlicht: (2026)
von: Papi, Sara, et al.
Veröffentlicht: (2026)
Good, but not always Fair: An Evaluation of Gender Bias for three commercial Machine Translation Systems
von: Piazzolla, Silvia Alma, et al.
Veröffentlicht: (2023)
von: Piazzolla, Silvia Alma, et al.
Veröffentlicht: (2023)
AlignAtt: Using Attention-based Audio-Translation Alignments as a Guide for Simultaneous Speech Translation
von: Papi, Sara, et al.
Veröffentlicht: (2023)
von: Papi, Sara, et al.
Veröffentlicht: (2023)
Prepending or Cross-Attention for Speech-to-Text? An Empirical Comparison
von: Lam, Tsz Kin, et al.
Veröffentlicht: (2025)
von: Lam, Tsz Kin, et al.
Veröffentlicht: (2025)
Speech-MASSIVE: A Multilingual Speech Dataset for SLU and Beyond
von: Lee, Beomseok, et al.
Veröffentlicht: (2024)
von: Lee, Beomseok, et al.
Veröffentlicht: (2024)
Speech Foundation Models and Crowdsourcing for Efficient, High-Quality Data Collection
von: Lee, Beomseok, et al.
Veröffentlicht: (2024)
von: Lee, Beomseok, et al.
Veröffentlicht: (2024)
Better Late Than Never: Meta-Evaluation of Latency Metrics for Simultaneous Speech-to-Text Translation
von: Polák, Peter, et al.
Veröffentlicht: (2025)
von: Polák, Peter, et al.
Veröffentlicht: (2025)
GAMBIT+: A Challenge Set for Evaluating Gender Bias in Machine Translation Quality Estimation Metrics
von: Filandrianos, Giorgos, et al.
Veröffentlicht: (2025)
von: Filandrianos, Giorgos, et al.
Veröffentlicht: (2025)
Instituto de Telecomunicações at IWSLT 2025: Aligning Small-Scale Speech and Language Models for Speech-to-Text Learning
von: Attanasio, Giuseppe, et al.
Veröffentlicht: (2025)
von: Attanasio, Giuseppe, et al.
Veröffentlicht: (2025)
When Good and Reproducible Results are a Giant with Feet of Clay: The Importance of Software Quality in NLP
von: Papi, Sara, et al.
Veröffentlicht: (2023)
von: Papi, Sara, et al.
Veröffentlicht: (2023)
NUTSHELL: A Dataset for Abstract Generation from Scientific Talks
von: Züfle, Maike, et al.
Veröffentlicht: (2025)
von: Züfle, Maike, et al.
Veröffentlicht: (2025)
GFG -- Gender-Fair Generation: A CALAMITA Challenge
von: Frenda, Simona, et al.
Veröffentlicht: (2024)
von: Frenda, Simona, et al.
Veröffentlicht: (2024)
Evaluating the IWSLT2023 Speech Translation Tasks: Human Annotations, Automatic Metrics, and Segmentation
von: Sperber, Matthias, et al.
Veröffentlicht: (2024)
von: Sperber, Matthias, et al.
Veröffentlicht: (2024)
Generative AI Practices, Literacy, and Divides: An Empirical Analysis in the Italian Context
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2025)
von: Savoldi, Beatrice, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Voice, Bias, and Coreference: An Interpretability Study of Gender in Speech Translation
von: Conti, Lina, et al.
Veröffentlicht: (2025) -
Speech Translation with Speech Foundation Models and Large Language Models: What is There and What is Missing?
von: Gaido, Marco, et al.
Veröffentlicht: (2024) -
Cross-Attention is Half Explanation in Speech-to-Text Models
von: Papi, Sara, et al.
Veröffentlicht: (2025) -
How do Hyenas deal with Human Speech? Speech Recognition and Translation with ConfHyena
von: Gaido, Marco, et al.
Veröffentlicht: (2024) -
SimulSeamless: FBK at IWSLT 2024 Simultaneous Speech Translation
von: Papi, Sara, et al.
Veröffentlicht: (2024)