Addressing speaker gender bias in large scale speech translation systems
Fuente:
arXiv
Salvato in:
| Autori principali: | Bansal, Shubham, Joshi, Vikas, Chadha, Harveen, Mehta, Rupeshkumar, Li, Jinyu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Length Aware Speech Translation for Video Dubbing
di: Chadha, Harveen Singh, et al.
Pubblicazione: (2025)
di: Chadha, Harveen Singh, et al.
Pubblicazione: (2025)
CTC-GMM: CTC guided modality matching for fast and accurate streaming speech translation
di: Zhao, Rui, et al.
Pubblicazione: (2024)
di: Zhao, Rui, et al.
Pubblicazione: (2024)
Language translation, and change of accent for speech-to-speech task using diffusion model
di: Mishra, Abhishek, et al.
Pubblicazione: (2025)
di: Mishra, Abhishek, et al.
Pubblicazione: (2025)
Are LLMs good pragmatic speakers?
di: Jian, Mingyue, et al.
Pubblicazione: (2024)
di: Jian, Mingyue, et al.
Pubblicazione: (2024)
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
di: Patel, Shubham, et al.
Pubblicazione: (2024)
di: Patel, Shubham, et al.
Pubblicazione: (2024)
Can we trust AI to detect healthy multilingual English speakers among the cognitively impaired cohort in the UK? An investigation using real-world conversational speech
di: Pahar, Madhurananda, et al.
Pubblicazione: (2026)
di: Pahar, Madhurananda, et al.
Pubblicazione: (2026)
Neural FOXP2 -- Language Specific Neuron Steering for Targeted Language Improvement in LLMs
di: Saha, Anusa, et al.
Pubblicazione: (2026)
di: Saha, Anusa, et al.
Pubblicazione: (2026)
XCB: an effective contextual biasing approach to bias cross-lingual phrases in speech recognition
di: Wan, Xucheng, et al.
Pubblicazione: (2024)
di: Wan, Xucheng, et al.
Pubblicazione: (2024)
AugSumm: towards generalizable speech summarization using synthetic labels from large language model
di: Jung, Jee-weon, et al.
Pubblicazione: (2024)
di: Jung, Jee-weon, et al.
Pubblicazione: (2024)
Generative Data Augmentation using LLMs improves Distributional Robustness in Question Answering
di: Chowdhury, Arijit Ghosh, et al.
Pubblicazione: (2023)
di: Chowdhury, Arijit Ghosh, et al.
Pubblicazione: (2023)
MindScope: Exploring cognitive biases in large language models through Multi-Agent Systems
di: Xie, Zhentao, et al.
Pubblicazione: (2024)
di: Xie, Zhentao, et al.
Pubblicazione: (2024)
Toward domain-specific machine translation and quality estimation systems
di: Sharami, Javad Pourmostafa Roshan
Pubblicazione: (2026)
di: Sharami, Javad Pourmostafa Roshan
Pubblicazione: (2026)
Inducing anxiety in large language models can induce bias
di: Coda-Forno, Julian, et al.
Pubblicazione: (2023)
di: Coda-Forno, Julian, et al.
Pubblicazione: (2023)
Simulating Meaning, Nevermore! Introducing ICR: A Semiotic-Hermeneutic Metric for Evaluating Meaning in LLM Text Summaries
di: Perez, Natalie, et al.
Pubblicazione: (2026)
di: Perez, Natalie, et al.
Pubblicazione: (2026)
A closer look at how large language models trust humans: patterns and biases
di: Lerman, Valeria, et al.
Pubblicazione: (2025)
di: Lerman, Valeria, et al.
Pubblicazione: (2025)
The order in speech disorder: a scoping review of state of the art machine learning methods for clinical speech classification
di: Moell, Birger, et al.
Pubblicazione: (2025)
di: Moell, Birger, et al.
Pubblicazione: (2025)
Defining bias in AI-systems: Biased models are fair models
di: Lindloff, Chiara, et al.
Pubblicazione: (2025)
di: Lindloff, Chiara, et al.
Pubblicazione: (2025)
Performance of a large language model-Artificial Intelligence based chatbot for counseling patients with sexually transmitted infections and genital diseases
di: Mehta, Nikhil, et al.
Pubblicazione: (2024)
di: Mehta, Nikhil, et al.
Pubblicazione: (2024)
Emergent effects of scaling on the functional hierarchies within large language models
di: Bogdan, Paul C.
Pubblicazione: (2025)
di: Bogdan, Paul C.
Pubblicazione: (2025)
LLMCache: Layer-Wise Caching Strategies for Accelerated Reuse in Transformer Inference
di: Bansal, Harsh Vardhan
Pubblicazione: (2025)
di: Bansal, Harsh Vardhan
Pubblicazione: (2025)
Human-Readable Adversarial Prompts: An Investigation into LLM Vulnerabilities Using Situational Context
di: Das, Nilanjana, et al.
Pubblicazione: (2024)
di: Das, Nilanjana, et al.
Pubblicazione: (2024)
Effective Reasoning Chains Reduce Intrinsic Dimensionality
di: Prasad, Archiki, et al.
Pubblicazione: (2026)
di: Prasad, Archiki, et al.
Pubblicazione: (2026)
PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training
di: Kumar, Harsh, et al.
Pubblicazione: (2026)
di: Kumar, Harsh, et al.
Pubblicazione: (2026)
Part-of-speech tagging for Nagamese Language using CRF
di: Shohe, Alovi N, et al.
Pubblicazione: (2025)
di: Shohe, Alovi N, et al.
Pubblicazione: (2025)
Tagarela - A Portuguese speech dataset from podcasts
di: de Oliveira, Frederico Santos, et al.
Pubblicazione: (2026)
di: de Oliveira, Frederico Santos, et al.
Pubblicazione: (2026)
Emotional Analysis of Fashion Trends Using Social Media and AI: Sentiment Analysis on Twitter for Fashion Trend Forecasting
di: Bansal, Aayam, et al.
Pubblicazione: (2025)
di: Bansal, Aayam, et al.
Pubblicazione: (2025)
Beyond the Parameters: A Technical Survey of Contextual Enrichment in Large Language Models: From In-Context Prompting to Causal Retrieval-Augmented Generation
di: Bansal, Prakhar, et al.
Pubblicazione: (2026)
di: Bansal, Prakhar, et al.
Pubblicazione: (2026)
Concept Attractors in LLMs and their Applications
di: Chytas, Sotirios Panagiotis, et al.
Pubblicazione: (2025)
di: Chytas, Sotirios Panagiotis, et al.
Pubblicazione: (2025)
A cross-species neural foundation model for end-to-end speech decoding
di: Zhang, Yizi, et al.
Pubblicazione: (2025)
di: Zhang, Yizi, et al.
Pubblicazione: (2025)
How good is GPT at writing political speeches for the White House?
di: Savoy, Jacques
Pubblicazione: (2024)
di: Savoy, Jacques
Pubblicazione: (2024)
Exposing and Addressing Cross-Task Inconsistency in Unified Vision-Language Models
di: Maharana, Adyasha, et al.
Pubblicazione: (2023)
di: Maharana, Adyasha, et al.
Pubblicazione: (2023)
A Survey of Prompt Engineering Methods in Large Language Models for Different NLP Tasks
di: Vatsal, Shubham, et al.
Pubblicazione: (2024)
di: Vatsal, Shubham, et al.
Pubblicazione: (2024)
Multilingual State Space Models for Structured Question Answering in Indic Languages
di: Vats, Arpita, et al.
Pubblicazione: (2025)
di: Vats, Arpita, et al.
Pubblicazione: (2025)
A Comprehensive Survey of Accelerated Generation Techniques in Large Language Models
di: Khoshnoodi, Mahsa, et al.
Pubblicazione: (2024)
di: Khoshnoodi, Mahsa, et al.
Pubblicazione: (2024)
Assessing LLM Reliability on Temporally Recent Open-Domain Questions
di: Krishnappa, Pushwitha, et al.
Pubblicazione: (2026)
di: Krishnappa, Pushwitha, et al.
Pubblicazione: (2026)
Reliability Analysis of Psychological Concept Extraction and Classification in User-penned Text
di: Garg, Muskan, et al.
Pubblicazione: (2024)
di: Garg, Muskan, et al.
Pubblicazione: (2024)
Evidence-backed Fact Checking using RAG and Few-Shot In-Context Learning with LLMs
di: Singhal, Ronit, et al.
Pubblicazione: (2024)
di: Singhal, Ronit, et al.
Pubblicazione: (2024)
Exploring the traditional NMT model and Large Language Model for chat translation
di: Yang, Jinlong, et al.
Pubblicazione: (2024)
di: Yang, Jinlong, et al.
Pubblicazione: (2024)
Mental Health Equity in LLMs: Leveraging Multi-Hop Question Answering to Detect Amplified and Silenced Perspectives
di: Haider, Batool, et al.
Pubblicazione: (2025)
di: Haider, Batool, et al.
Pubblicazione: (2025)
Born With a Silver Spoon? Investigating Socioeconomic Bias in Large Language Models
di: Singh, Smriti, et al.
Pubblicazione: (2024)
di: Singh, Smriti, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Length Aware Speech Translation for Video Dubbing
di: Chadha, Harveen Singh, et al.
Pubblicazione: (2025) -
CTC-GMM: CTC guided modality matching for fast and accurate streaming speech translation
di: Zhao, Rui, et al.
Pubblicazione: (2024) -
Language translation, and change of accent for speech-to-speech task using diffusion model
di: Mishra, Abhishek, et al.
Pubblicazione: (2025) -
Are LLMs good pragmatic speakers?
di: Jian, Mingyue, et al.
Pubblicazione: (2024) -
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
di: Patel, Shubham, et al.
Pubblicazione: (2024)