Edresson Casanova
Featured Co-authors
- David Ifeoluwa Adelani 25 publications
- Elizabeth Salesky 22 publications
- Iroro Orife 15 publications
- Salomey Osei 14 publications
- Marcelo Finger 13 publications
- Arnaldo Candido Junior 12 publications
- Moacir Antonelli Ponti 12 publications
- Sandra Maria Aluisio 11 publications
- Chris Emezue 9 publications
- Jesujoba Alabi 9 publications
- Perez Ogayo 9 publications
Research Publications
Evaluation of Speech Representations for MOS prediction
In this paper, we evaluate feature extraction models for predicting speech...
Authors: Frederico S. Oliveira, et al.
CML-TTS A Multilingual Dataset for Speech Synthesis in Low-Resource Languages
In this paper, we present CML-TTS, a recursive acronym for CML-Multi-Lin...
Authors: Frederico S. Oliveira, et al.
Evaluating OpenAI's Whisper ASR for Punctuation Prediction and Topic Modeling of life histories of the Museum of the Person
Automatic speech recognition (ASR) systems play a key role in applications...
Authors: Lucas Rafael Stefanel Gris, et al.
Interpretability Analysis of Deep Models for COVID-19 Detection
During the outbreak of the COVID-19 pandemic, several research areas joined...
Authors: Daniel Peixoto Pinto da Silva, et al.
BibleTTS: a large, high-fidelity, multilingual, and uniquely African speech corpus
BibleTTS is a large, high-quality, open speech dataset for ten languages...
Authors: Josh Meyer, et al.
A single speaker is almost all you need for automatic speech recognition
We explore the use of speech synthesis and voice conversion applied to...
Authors: Edresson Casanova, et al.
YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone
YourTTS brings the power of a multilingual approach to the task of zero...
Authors: Edresson Casanova, et al.
CORAA: a large corpus of spontaneous and prepared speech manually validated for speech recognition in Brazilian Portuguese
Automatic Speech recognition (ASR) is a complex and challenging task. In...
Authors: Arnaldo Candido Junior, et al.
Brazilian Portuguese Speech Recognition Using Wav2vec 2.0
Deep learning techniques have been shown to be efficient in various tasks...
Authors: Lucas Rafael Stefanel Gris, et al.
SC-GlowTTS: an Efficient Zero-Shot Multi-Speaker Text-To-Speech Model
In this paper, we propose SC-GlowTTS: an efficient zero-shot multi-speak...
Authors: Edresson Casanova, et al.
End-To-End Speech Synthesis Applied to Brazilian Portuguese
Voice synthesis systems are popular in different applications, such as...
Authors: Edresson Casanova, et al.
Speech2Phone: A Multilingual and Text Independent Speaker Identification Model
Voice recognition is an area with a wide application potential. Speaker...
Authors: Edresson Casanova, et al.