Skip to content

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

1 Commit
 
 

Repository files navigation

Tamil-Speech-Dataset-500-Hours-Monologue-Audio-Corpus

Description

This dataset includes 500 hours of scripted Tamil monologue speech collected using smartphones. Each sample is transcribed with text content and metadata such as speaker ID, gender, and age. The dataset features diverse speakers from various regions, making it highly representative of real-world Tamil language use and suitable for automatic speech recognition (ASR), text-to-speech (TTS), voice activity detection (VAD), and natural language processing (NLP) tasks

For more details, please refer to the link: https://www.nexdata.ai/datasets/speechrecog/1838?source=Github

Format

16kHz, 16bit, uncompressed wav, mono channel.

Recording condition

quiet indoor environment, low background noise, without echo;

Recording device

Android smartphone, iPhone;

Speaker

479 speakers totally, with 52% female and 48% male

Language

Tamil;

Features of annotation

Transcription text;

Accuracy Rate

Word Accuracy Rate (WAR) 95%;

Licensing Information

Commercial License

About

No description, website, or topics provided.

Resources

Stars

Watchers

Forks

Releases

Packages

Contributors