whisply: Cross-Platform Python App for Batch Transcription, Translation, Speaker Annotation and Subtitle Generation of Video and Audio Content
Fuente:
Zenodo
Enregistré dans:
| Auteur principal: | Schmidt, Thomas |
|---|---|
| Format: | Recurso digital |
| Publié: |
Zenodo
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Ep. 7: Building Custom ASR Tools
par: Rosehill, Daniel, et autres
Publié: (2025)
par: Rosehill, Daniel, et autres
Publié: (2025)
Communication Interface for Mexican Spanish Dysarthric Speakers.
par: Gladys Bonilla-Enriquez
Publié: (2012)
par: Gladys Bonilla-Enriquez
Publié: (2012)
Category-based Language Models in a Spanish Spoken Dialogue System
par: Raquel Justo
Publié: (2006)
par: Raquel Justo
Publié: (2006)
The AI Transcription Sweet Spot
par: Rosehill, Daniel, et autres
Publié: (2026)
par: Rosehill, Daniel, et autres
Publié: (2026)
3D Modeling of the Mexican Sign Language for a Speech-to-Sign Language System
par: Santiago-Omar Caballero-Morales
Publié: (2013)
par: Santiago-Omar Caballero-Morales
Publié: (2013)
Robust Spoken Language Understanding for House Service Robots
par: Andrea Vanzo
Publié: (2016)
par: Andrea Vanzo
Publié: (2016)
What benefits can speech recognition and ambient AI software provide to the healthcare sector?
par: Tripdatabase
Publié: (2025)
par: Tripdatabase
Publié: (2025)
Enhacement of the Imput Interface of Spoken Dialogue System By Means of Contextual Models and Grammatical Rules
par: Ramón López-Cózar
Publié: (2009)
par: Ramón López-Cózar
Publié: (2009)
Enrolment-based personalisation for improving individual-level fairness in speech emotion recognition
par: Triantafyllopoulos, Andreas, et autres
Publié: (2024)
par: Triantafyllopoulos, Andreas, et autres
Publié: (2024)
A New and Efficient Alignment Technique by Cosine Distance
par: Alain Manzo-Martínez
Publié: (2013)
par: Alain Manzo-Martínez
Publié: (2013)
ASLP-MULAN: Audio speech and language processing for multimedia analytics
par: Javier Ferreiros
Publié: (2016)
par: Javier Ferreiros
Publié: (2016)
Ep. 15: AI Gets Personal: The Power of Voice Fine-Tuning
par: Rosehill, Daniel, et autres
Publié: (2025)
par: Rosehill, Daniel, et autres
Publié: (2025)
The Voice Keyboard: Killing the "Digital Sandwich"
par: Rosehill, Daniel, et autres
Publié: (2026)
par: Rosehill, Daniel, et autres
Publié: (2026)
Automatic speech recognizers for Mexican Spanish and its open resources
par: Carlos Daniel Hernández-Mena
Publié: (2017)
par: Carlos Daniel Hernández-Mena
Publié: (2017)
Reconocedor de Palabras con el uso de Regresión Lineal y Coeficiente Muestral
par: Carlos Alejandro de Luna Ortega
Publié: (2012)
par: Carlos Alejandro de Luna Ortega
Publié: (2012)
HOW WELL CAN ASR TECHNOLOGY UNDERSTAND FOREIGN-ACCENTED SPEECH?
par: Souza, Hanna Kivistö, et autres
Publié: (2022)
par: Souza, Hanna Kivistö, et autres
Publié: (2022)
Automatic prediction of emotions from text in Spanish for expressive speech synthesis in the chat domain
par: Benjamin Kolz
Publié: (2014)
par: Benjamin Kolz
Publié: (2014)
Towards automatic recognition of irregular, short-open answers in Fill-in-the-blank tests
par: SERGIO A. ROJAS
Publié: (2014)
par: SERGIO A. ROJAS
Publié: (2014)
Profiling Hate Speech Spreaders on Twitter
par: FRANCISCO RANGEL, et autres
Publié: (2021)
par: FRANCISCO RANGEL, et autres
Publié: (2021)
Arguing as Trying to Show That a Target-claim is Correct
par: DAVID HITCHCOCK
Publié: (2011)
par: DAVID HITCHCOCK
Publié: (2011)
NORMS OF TEACHER'S SPEECH
par: Ahmadova Malika Feruz qizi
Publié: (2026)
par: Ahmadova Malika Feruz qizi
Publié: (2026)
Exploratory Study on the Prevalence of Speech Sound Disorders in a Group of Valencian School Students Belonging to 3rd Grade of Infant School and 1st Grade of Primary School
par: Omaya Amr Rey
Publié: (2022)
par: Omaya Amr Rey
Publié: (2022)
Nitq aktının psixolinqvistik xüsusiyyətləri
par: Həbibova, Könül
Publié: (2019)
par: Həbibova, Könül
Publié: (2019)
Proceedings of the Scientific-Practical Conference "Research and Development - 2016"
Publié: (2018)
Publié: (2018)
Психолінгвістика
Publié: (2018)
Publié: (2018)
Efficacy of the elementary, middle, and high school students' persuasive speech: Evidence from South Korea
par: Kim, Yune Jung, et autres
Publié: (2023)
par: Kim, Yune Jung, et autres
Publié: (2023)
A pattern recognition based esophageal speech enhancement system
par: A. Mantilla-Caeiros
Publié: (2010)
par: A. Mantilla-Caeiros
Publié: (2010)
EURASIP Journal on Audio, Speech, and Music Processing
Publié: (2006)
Publié: (2006)
Dil, Konuşma ve Yutma Araştırmaları Dergisi
Publié: (2023)
Publié: (2023)
Word final prolongations: acoustic characteristics and influence on speech fluency perception
par: Souza, Lívia Maria Santos de, et autres
Publié: (2023)
par: Souza, Lívia Maria Santos de, et autres
Publié: (2023)
DÜNYANIN DILSEL MANZARASI VE TÜRKÇEDEKI SÖZCÜK TÜRLERI MESELESI
par: ELTAZAROV, Jo'liboy
Publié: (2025)
par: ELTAZAROV, Jo'liboy
Publié: (2025)
Logopedia
Publié: (2024)
Publié: (2024)
Dil ve Konuşma Terapisinin İnterdisipliner ve Multidisipliner Alandaki Farkındalığı Çalışması
par: Murat Berke Kurt, et autres
Publié: (2019)
par: Murat Berke Kurt, et autres
Publié: (2019)
Jurnal Linguistik Komputasional
Publié: (2019)
Publié: (2019)
Audiology: Communication Research
Publié: (2016)
Publié: (2016)
Logopedia Silesiana
Publié: (2022)
Publié: (2022)
Chatterbox TTS: Open Source vs. ElevenLabs
par: Rosehill, Daniel, et autres
Publié: (2026)
par: Rosehill, Daniel, et autres
Publié: (2026)
Dialogical surface text features in abstracts
par: Ingrid García-Østbye
Publié: (2008)
par: Ingrid García-Østbye
Publié: (2008)
Software development to help individuals with communication limitations or complete absence of speech: a systematic review
par: de Paula Silva, Fernanda, et autres
Publié: (2026)
par: de Paula Silva, Fernanda, et autres
Publié: (2026)
Rapaloid: Adapting an existing framework for emotional speech prosody generation
par: Marc Freixes
Publié: (2012)
par: Marc Freixes
Publié: (2012)
Documents similaires
-
Ep. 7: Building Custom ASR Tools
par: Rosehill, Daniel, et autres
Publié: (2025) -
Communication Interface for Mexican Spanish Dysarthric Speakers.
par: Gladys Bonilla-Enriquez
Publié: (2012) -
Category-based Language Models in a Spanish Spoken Dialogue System
par: Raquel Justo
Publié: (2006) -
The AI Transcription Sweet Spot
par: Rosehill, Daniel, et autres
Publié: (2026) -
3D Modeling of the Mexican Sign Language for a Speech-to-Sign Language System
par: Santiago-Omar Caballero-Morales
Publié: (2013)