Skip to content

Latest commit

 

History

History
24 lines (20 loc) · 943 Bytes

File metadata and controls

24 lines (20 loc) · 943 Bytes

496-Hours-Hindi-Speech-Data-by-Mobile-Phone

Description

Hindi(India) Scripted Monologue Smartphone speech dataset, collected from monologue based on given scripts. Transcribed with text content and other attributes. Our dataset was collected from extensive and diversify speakers(551 Indian native speakers), geographicly speaking, enhancing model performance in real and complex tasks.

For more details, please refer to the link: https://www.nexdata.ai/datasets/speechrecog/1264?source=Github

Format

16kHz, 16bit, uncompressed wav, mono channel.

Recording environment

quiet indoor environment, low background noise, without echo.

Recording content (read speech)

news and oral category;

Demographics

551 speakers totally, with 43% males and 57% females;

Device

Android mobile phone, iPhone

Language

Hindi

Application scenarios

speech recognition; voiceprint recognition

Licensing Information

Commercial License