Creating an African American-Sounding TTS: Guidelines, Technical Challenges,and Surprising Evaluations
Fuente:
arXiv
Saved in:
| Main Authors: | Pinhanez, Claudio, Fernandez, Raul, Grave, Marcelo, Nogima, Julio, Hoory, Ron |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating the Usage of African-American Vernacular English in Large Language Models
by: Dunlap, Deja, et al.
Published: (2026)
by: Dunlap, Deja, et al.
Published: (2026)
Creating and Evaluating Personas Using Generative AI: A Scoping Review of 81 Articles
by: Amin, Danial, et al.
Published: (2025)
by: Amin, Danial, et al.
Published: (2025)
Towards a Design Guideline for RPA Evaluation: A Survey of Large Language Model-Based Role-Playing Agents
by: Chen, Chaoran, et al.
Published: (2025)
by: Chen, Chaoran, et al.
Published: (2025)
Six Guidelines for Trustworthy, Ethical and Responsible Automation Design
by: Jelínek, Matouš, et al.
Published: (2025)
by: Jelínek, Matouš, et al.
Published: (2025)
Advancing Social Intelligence in AI Agents: Technical Challenges and Open Questions
by: Mathur, Leena, et al.
Published: (2024)
by: Mathur, Leena, et al.
Published: (2024)
The Reader is the Metric: How Textual Features and Reader Profiles Explain Conflicting Evaluations of AI Creative Writing
by: Marco, Guillermo, et al.
Published: (2025)
by: Marco, Guillermo, et al.
Published: (2025)
Aligning Stuttered-Speech Research with End-User Needs: Scoping Review, Survey, and Guidelines
by: Toyin, Hawau Olamide, et al.
Published: (2026)
by: Toyin, Hawau Olamide, et al.
Published: (2026)
Using Generative Text Models to Create Qualitative Codebooks for Student Evaluations of Teaching
by: Katz, Andrew, et al.
Published: (2024)
by: Katz, Andrew, et al.
Published: (2024)
Using a Human-AI Teaming Approach to Create and Curate Scientific Datasets with the SCILIRE System
by: Bölücü, Necva, et al.
Published: (2026)
by: Bölücü, Necva, et al.
Published: (2026)
Evaluating Generative AI in the Lab: Methodological Challenges and Guidelines
by: Park, Hyerim, et al.
Published: (2026)
by: Park, Hyerim, et al.
Published: (2026)
Roleplay-doh: Enabling Domain-Experts to Create LLM-simulated Patients via Eliciting and Adhering to Principles
by: Louie, Ryan, et al.
Published: (2024)
by: Louie, Ryan, et al.
Published: (2024)
RAI Guidelines: Method for Generating Responsible AI Guidelines Grounded in Regulations and Usable by (Non-)Technical Roles
by: Constantinides, Marios, et al.
Published: (2023)
by: Constantinides, Marios, et al.
Published: (2023)
Creating General User Models from Computer Use
by: Shaikh, Omar, et al.
Published: (2025)
by: Shaikh, Omar, et al.
Published: (2025)
AI as a deliberative partner fosters intercultural empathy for Americans but fails for Latin American participants
by: Villanueva, Isabel, et al.
Published: (2025)
by: Villanueva, Isabel, et al.
Published: (2025)
Revealing and Mitigating the Challenge of Detecting Character Knowledge Errors in LLM Role-Playing
by: Zhang, Wenyuan, et al.
Published: (2024)
by: Zhang, Wenyuan, et al.
Published: (2024)
Exploring Human Perceptions of AI Responses: Insights from a Mixed-Methods Study on Risk Mitigation in Generative Models
by: Candello, Heloisa, et al.
Published: (2025)
by: Candello, Heloisa, et al.
Published: (2025)
Word Synchronization Challenge: A Benchmark for Word Association Responses for Large Language Models
by: Cazalets, Tanguy, et al.
Published: (2025)
by: Cazalets, Tanguy, et al.
Published: (2025)
Understanding Learner-LLM Chatbot Interactions and the Impact of Prompting Guidelines
by: Koyuturk, Cansu, et al.
Published: (2025)
by: Koyuturk, Cansu, et al.
Published: (2025)
CNER: A tool Classifier of Named-Entity Relationships
by: Torres, Jefferson A. Peña, et al.
Published: (2024)
by: Torres, Jefferson A. Peña, et al.
Published: (2024)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
by: Vo, Truong, et al.
Published: (2025)
by: Vo, Truong, et al.
Published: (2025)
WAXAL-NET: Finetuned Edge ASR Across 19 African Languages
by: Olufemi, Victor Tolulope, et al.
Published: (2026)
by: Olufemi, Victor Tolulope, et al.
Published: (2026)
Your Co-Workers Matter: Evaluating Collaborative Capabilities of Language Models in Blocks World
by: Wu, Guande, et al.
Published: (2024)
by: Wu, Guande, et al.
Published: (2024)
Red Teaming LLMs as Socio-Technical Practice: From Exploration and Data Creation to Evaluation
by: Garcia, Adriana Alvarado, et al.
Published: (2026)
by: Garcia, Adriana Alvarado, et al.
Published: (2026)
Robots in the Middle: Evaluating LLMs in Dispute Resolution
by: Tan, Jinzhe, et al.
Published: (2024)
by: Tan, Jinzhe, et al.
Published: (2024)
Pearmut: Human Evaluation of Translation Made Trivial
by: Zouhar, Vilém, et al.
Published: (2026)
by: Zouhar, Vilém, et al.
Published: (2026)
Conversations Gone Awry, But Then? Evaluating Conversational Forecasting Models
by: Tran, Son Quoc, et al.
Published: (2025)
by: Tran, Son Quoc, et al.
Published: (2025)
Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations
by: Liang, Chen, et al.
Published: (2026)
by: Liang, Chen, et al.
Published: (2026)
Context-Aware Monolingual Human Evaluation of Machine Translation
by: Picinini, Silvio, et al.
Published: (2025)
by: Picinini, Silvio, et al.
Published: (2025)
Designing and Evaluating Chain-of-Hints for Scientific Question Answering
by: Jangra, Anubhav, et al.
Published: (2025)
by: Jangra, Anubhav, et al.
Published: (2025)
ELI-Why: Evaluating the Pedagogical Utility of Language Model Explanations
by: Joshi, Brihi, et al.
Published: (2025)
by: Joshi, Brihi, et al.
Published: (2025)
DICE: A Framework for Dimensional and Contextual Evaluation of Language Models
by: Shrivastava, Aryan, et al.
Published: (2025)
by: Shrivastava, Aryan, et al.
Published: (2025)
Audio-Based Crowd-Sourced Evaluation of Machine Translation Quality
by: Haq, Sami Ul, et al.
Published: (2025)
by: Haq, Sami Ul, et al.
Published: (2025)
Detection and Positive Reconstruction of Cognitive Distortion sentences: Mandarin Dataset and Evaluation
by: Lin, Shuya, et al.
Published: (2024)
by: Lin, Shuya, et al.
Published: (2024)
PALLM: Evaluating and Enhancing PALLiative Care Conversations with Large Language Models
by: Wang, Zhiyuan, et al.
Published: (2024)
by: Wang, Zhiyuan, et al.
Published: (2024)
Evaluation of a Sign Language Avatar on Comprehensibility, User Experience \& Acceptability
by: Wasserroth, Fenya, et al.
Published: (2025)
by: Wasserroth, Fenya, et al.
Published: (2025)
IP-Dialog: Evaluating Implicit Personalization in Dialogue Systems with Synthetic Data
by: Peng, Bo, et al.
Published: (2025)
by: Peng, Bo, et al.
Published: (2025)
HEDS 3.0: The Human Evaluation Data Sheet Version 3.0
by: Belz, Anya, et al.
Published: (2024)
by: Belz, Anya, et al.
Published: (2024)
Evaluating LLM-Generated Q&A Test: a Student-Centered Study
by: Wróblewska, Anna, et al.
Published: (2025)
by: Wróblewska, Anna, et al.
Published: (2025)
EmoHarbor: Evaluating Personalized Emotional Support by Simulating the User's Internal World
by: Ye, Jing, et al.
Published: (2026)
by: Ye, Jing, et al.
Published: (2026)
HammerBench: Fine-Grained Function-Calling Evaluation in Real Mobile Device Scenarios
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
Similar Items
-
Evaluating the Usage of African-American Vernacular English in Large Language Models
by: Dunlap, Deja, et al.
Published: (2026) -
Creating and Evaluating Personas Using Generative AI: A Scoping Review of 81 Articles
by: Amin, Danial, et al.
Published: (2025) -
Towards a Design Guideline for RPA Evaluation: A Survey of Large Language Model-Based Role-Playing Agents
by: Chen, Chaoran, et al.
Published: (2025) -
Six Guidelines for Trustworthy, Ethical and Responsible Automation Design
by: Jelínek, Matouš, et al.
Published: (2025) -
Advancing Social Intelligence in AI Agents: Technical Challenges and Open Questions
by: Mathur, Leena, et al.
Published: (2024)