Skip to content

Nexdata-AI/10-Hour-Brazilian-Portuguese-TTS-Dataset

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

1 Commit
 
 

Repository files navigation

10-Hour-Brazilian-Portuguese-TTS-Dataset

Description

This dataset contains 10 hours of Brazilian Portuguese speech recordings, collected from native Brazilian speakers. The corpus is related to the customer service field. The dataset features balanced phoneme coverage. Professional phonetician participates in the annotation. It precisely matches with the research and development needs of the speech synthesis.

For more details, please refer to the link: https://www.nexdata.ai/datasets/tts/1895?source=Github

Format

48,000Hz, 24bit, uncompressed wav, mono channel;

Recording environment

professional recording studio;

Recording content

customer service;

Speaker

Brazilian;

Annotation

word and phoneme transcription, prosodic boundary annotation;

Device

microphone;

Language

Brazilian Portuguese;

Application scenarios

speech synthesis.

Licensing Information

Commercial License

About

No description, website, or topics provided.

Resources

Stars

Watchers

Forks

Releases

Packages

Contributors