Addressing speaker gender bias in large scale speech translation systems
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bansal, Shubham, Joshi, Vikas, Chadha, Harveen, Mehta, Rupeshkumar, Li, Jinyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Length Aware Speech Translation for Video Dubbing
von: Chadha, Harveen Singh, et al.
Veröffentlicht: (2025)
von: Chadha, Harveen Singh, et al.
Veröffentlicht: (2025)
CTC-GMM: CTC guided modality matching for fast and accurate streaming speech translation
von: Zhao, Rui, et al.
Veröffentlicht: (2024)
von: Zhao, Rui, et al.
Veröffentlicht: (2024)
Language translation, and change of accent for speech-to-speech task using diffusion model
von: Mishra, Abhishek, et al.
Veröffentlicht: (2025)
von: Mishra, Abhishek, et al.
Veröffentlicht: (2025)
Are LLMs good pragmatic speakers?
von: Jian, Mingyue, et al.
Veröffentlicht: (2024)
von: Jian, Mingyue, et al.
Veröffentlicht: (2024)
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
von: Patel, Shubham, et al.
Veröffentlicht: (2024)
von: Patel, Shubham, et al.
Veröffentlicht: (2024)
Can we trust AI to detect healthy multilingual English speakers among the cognitively impaired cohort in the UK? An investigation using real-world conversational speech
von: Pahar, Madhurananda, et al.
Veröffentlicht: (2026)
von: Pahar, Madhurananda, et al.
Veröffentlicht: (2026)
Neural FOXP2 -- Language Specific Neuron Steering for Targeted Language Improvement in LLMs
von: Saha, Anusa, et al.
Veröffentlicht: (2026)
von: Saha, Anusa, et al.
Veröffentlicht: (2026)
XCB: an effective contextual biasing approach to bias cross-lingual phrases in speech recognition
von: Wan, Xucheng, et al.
Veröffentlicht: (2024)
von: Wan, Xucheng, et al.
Veröffentlicht: (2024)
AugSumm: towards generalizable speech summarization using synthetic labels from large language model
von: Jung, Jee-weon, et al.
Veröffentlicht: (2024)
von: Jung, Jee-weon, et al.
Veröffentlicht: (2024)
Generative Data Augmentation using LLMs improves Distributional Robustness in Question Answering
von: Chowdhury, Arijit Ghosh, et al.
Veröffentlicht: (2023)
von: Chowdhury, Arijit Ghosh, et al.
Veröffentlicht: (2023)
MindScope: Exploring cognitive biases in large language models through Multi-Agent Systems
von: Xie, Zhentao, et al.
Veröffentlicht: (2024)
von: Xie, Zhentao, et al.
Veröffentlicht: (2024)
Toward domain-specific machine translation and quality estimation systems
von: Sharami, Javad Pourmostafa Roshan
Veröffentlicht: (2026)
von: Sharami, Javad Pourmostafa Roshan
Veröffentlicht: (2026)
Inducing anxiety in large language models can induce bias
von: Coda-Forno, Julian, et al.
Veröffentlicht: (2023)
von: Coda-Forno, Julian, et al.
Veröffentlicht: (2023)
Simulating Meaning, Nevermore! Introducing ICR: A Semiotic-Hermeneutic Metric for Evaluating Meaning in LLM Text Summaries
von: Perez, Natalie, et al.
Veröffentlicht: (2026)
von: Perez, Natalie, et al.
Veröffentlicht: (2026)
A closer look at how large language models trust humans: patterns and biases
von: Lerman, Valeria, et al.
Veröffentlicht: (2025)
von: Lerman, Valeria, et al.
Veröffentlicht: (2025)
The order in speech disorder: a scoping review of state of the art machine learning methods for clinical speech classification
von: Moell, Birger, et al.
Veröffentlicht: (2025)
von: Moell, Birger, et al.
Veröffentlicht: (2025)
Defining bias in AI-systems: Biased models are fair models
von: Lindloff, Chiara, et al.
Veröffentlicht: (2025)
von: Lindloff, Chiara, et al.
Veröffentlicht: (2025)
Performance of a large language model-Artificial Intelligence based chatbot for counseling patients with sexually transmitted infections and genital diseases
von: Mehta, Nikhil, et al.
Veröffentlicht: (2024)
von: Mehta, Nikhil, et al.
Veröffentlicht: (2024)
Emergent effects of scaling on the functional hierarchies within large language models
von: Bogdan, Paul C.
Veröffentlicht: (2025)
von: Bogdan, Paul C.
Veröffentlicht: (2025)
LLMCache: Layer-Wise Caching Strategies for Accelerated Reuse in Transformer Inference
von: Bansal, Harsh Vardhan
Veröffentlicht: (2025)
von: Bansal, Harsh Vardhan
Veröffentlicht: (2025)
Human-Readable Adversarial Prompts: An Investigation into LLM Vulnerabilities Using Situational Context
von: Das, Nilanjana, et al.
Veröffentlicht: (2024)
von: Das, Nilanjana, et al.
Veröffentlicht: (2024)
Effective Reasoning Chains Reduce Intrinsic Dimensionality
von: Prasad, Archiki, et al.
Veröffentlicht: (2026)
von: Prasad, Archiki, et al.
Veröffentlicht: (2026)
PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Training
von: Kumar, Harsh, et al.
Veröffentlicht: (2026)
von: Kumar, Harsh, et al.
Veröffentlicht: (2026)
Part-of-speech tagging for Nagamese Language using CRF
von: Shohe, Alovi N, et al.
Veröffentlicht: (2025)
von: Shohe, Alovi N, et al.
Veröffentlicht: (2025)
Tagarela - A Portuguese speech dataset from podcasts
von: de Oliveira, Frederico Santos, et al.
Veröffentlicht: (2026)
von: de Oliveira, Frederico Santos, et al.
Veröffentlicht: (2026)
Emotional Analysis of Fashion Trends Using Social Media and AI: Sentiment Analysis on Twitter for Fashion Trend Forecasting
von: Bansal, Aayam, et al.
Veröffentlicht: (2025)
von: Bansal, Aayam, et al.
Veröffentlicht: (2025)
Beyond the Parameters: A Technical Survey of Contextual Enrichment in Large Language Models: From In-Context Prompting to Causal Retrieval-Augmented Generation
von: Bansal, Prakhar, et al.
Veröffentlicht: (2026)
von: Bansal, Prakhar, et al.
Veröffentlicht: (2026)
Concept Attractors in LLMs and their Applications
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2025)
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2025)
A cross-species neural foundation model for end-to-end speech decoding
von: Zhang, Yizi, et al.
Veröffentlicht: (2025)
von: Zhang, Yizi, et al.
Veröffentlicht: (2025)
How good is GPT at writing political speeches for the White House?
von: Savoy, Jacques
Veröffentlicht: (2024)
von: Savoy, Jacques
Veröffentlicht: (2024)
Exposing and Addressing Cross-Task Inconsistency in Unified Vision-Language Models
von: Maharana, Adyasha, et al.
Veröffentlicht: (2023)
von: Maharana, Adyasha, et al.
Veröffentlicht: (2023)
A Survey of Prompt Engineering Methods in Large Language Models for Different NLP Tasks
von: Vatsal, Shubham, et al.
Veröffentlicht: (2024)
von: Vatsal, Shubham, et al.
Veröffentlicht: (2024)
Multilingual State Space Models for Structured Question Answering in Indic Languages
von: Vats, Arpita, et al.
Veröffentlicht: (2025)
von: Vats, Arpita, et al.
Veröffentlicht: (2025)
A Comprehensive Survey of Accelerated Generation Techniques in Large Language Models
von: Khoshnoodi, Mahsa, et al.
Veröffentlicht: (2024)
von: Khoshnoodi, Mahsa, et al.
Veröffentlicht: (2024)
Assessing LLM Reliability on Temporally Recent Open-Domain Questions
von: Krishnappa, Pushwitha, et al.
Veröffentlicht: (2026)
von: Krishnappa, Pushwitha, et al.
Veröffentlicht: (2026)
Reliability Analysis of Psychological Concept Extraction and Classification in User-penned Text
von: Garg, Muskan, et al.
Veröffentlicht: (2024)
von: Garg, Muskan, et al.
Veröffentlicht: (2024)
Evidence-backed Fact Checking using RAG and Few-Shot In-Context Learning with LLMs
von: Singhal, Ronit, et al.
Veröffentlicht: (2024)
von: Singhal, Ronit, et al.
Veröffentlicht: (2024)
Exploring the traditional NMT model and Large Language Model for chat translation
von: Yang, Jinlong, et al.
Veröffentlicht: (2024)
von: Yang, Jinlong, et al.
Veröffentlicht: (2024)
Mental Health Equity in LLMs: Leveraging Multi-Hop Question Answering to Detect Amplified and Silenced Perspectives
von: Haider, Batool, et al.
Veröffentlicht: (2025)
von: Haider, Batool, et al.
Veröffentlicht: (2025)
Born With a Silver Spoon? Investigating Socioeconomic Bias in Large Language Models
von: Singh, Smriti, et al.
Veröffentlicht: (2024)
von: Singh, Smriti, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Length Aware Speech Translation for Video Dubbing
von: Chadha, Harveen Singh, et al.
Veröffentlicht: (2025) -
CTC-GMM: CTC guided modality matching for fast and accurate streaming speech translation
von: Zhao, Rui, et al.
Veröffentlicht: (2024) -
Language translation, and change of accent for speech-to-speech task using diffusion model
von: Mishra, Abhishek, et al.
Veröffentlicht: (2025) -
Are LLMs good pragmatic speakers?
von: Jian, Mingyue, et al.
Veröffentlicht: (2024) -
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
von: Patel, Shubham, et al.
Veröffentlicht: (2024)