Robat-e-Beheshti: A Persian Wake Word Detection Dataset for Robotic Purposes
Parisa Ahmadzadeh Raji, Yasser Shekofteh
- 发表年份
- 2022
- 引用次数
- 2
摘要
In this paper a dataset is explained which was collected for a wake word detection project which works on classifying audio data. The data are audio files in Persian language. The number of 5738 audio files with different formats and a maximum length of 3 seconds were collected. Afterwards the format of audio files and their sample rates were changed to a specific one. There are positive and negative examples for this dataset. Different audio recorders were used for recording the audios and most of the data were gathered from 187 individuals and the other files were collected from the open source ShEMO dataset: a large-scale validated database for Persian speech emotion detection. In this paper the process of collecting data and how they were organized in different files, are explained and the data analysis is also presented. This dataset will be free for academic usage with official request to the authors.
关键词
相关论文
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
1999
Artificial intelligence: a modern approach
1995
Fractional Differential Equations
Igor Podlubný
2025
Applied Nonlinear Control
Jean-Jacques Slotine, Weiping Li
1991