Edresson Casanova

Featured Co-authors

Research Publications

Evaluation of Speech Representations for MOS prediction

In this paper, we evaluate feature extraction models for predicting speech...

Authors: Frederico S. Oliveira, et al.

CML-TTS A Multilingual Dataset for Speech Synthesis in Low-Resource Languages

In this paper, we present CML-TTS, a recursive acronym for CML-Multi-Lin...

Authors: Frederico S. Oliveira, et al.

Evaluating OpenAI's Whisper ASR for Punctuation Prediction and Topic Modeling of life histories of the Museum of the Person

Automatic speech recognition (ASR) systems play a key role in applications...

Authors: Lucas Rafael Stefanel Gris, et al.

Interpretability Analysis of Deep Models for COVID-19 Detection

During the outbreak of the COVID-19 pandemic, several research areas joined...

Authors: Daniel Peixoto Pinto da Silva, et al.

BibleTTS: a large, high-fidelity, multilingual, and uniquely African speech corpus

BibleTTS is a large, high-quality, open speech dataset for ten languages...

Authors: Josh Meyer, et al.

A single speaker is almost all you need for automatic speech recognition

We explore the use of speech synthesis and voice conversion applied to...

Authors: Edresson Casanova, et al.

YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone

YourTTS brings the power of a multilingual approach to the task of zero...

Authors: Edresson Casanova, et al.

CORAA: a large corpus of spontaneous and prepared speech manually validated for speech recognition in Brazilian Portuguese

Automatic Speech recognition (ASR) is a complex and challenging task. In...

Authors: Arnaldo Candido Junior, et al.

Brazilian Portuguese Speech Recognition Using Wav2vec 2.0

Deep learning techniques have been shown to be efficient in various tasks...

Authors: Lucas Rafael Stefanel Gris, et al.

SC-GlowTTS: an Efficient Zero-Shot Multi-Speaker Text-To-Speech Model

In this paper, we propose SC-GlowTTS: an efficient zero-shot multi-speak...

Authors: Edresson Casanova, et al.

End-To-End Speech Synthesis Applied to Brazilian Portuguese

Voice synthesis systems are popular in different applications, such as...

Authors: Edresson Casanova, et al.

Speech2Phone: A Multilingual and Text Independent Speaker Identification Model

Voice recognition is an area with a wide application potential. Speaker...

Authors: Edresson Casanova, et al.