Content deleted Content added
m Bot: Removing category Category:Artificial intelligence, bcoz it's already in Category:Applications of artificial intelligence |
|||
(29 intermediate revisions by 9 users not shown) | |||
Line 1:
The '''PCVC (Persian Consonant Vowel Combination) Speech Dataset''' is a [[Modern Persian]] [[speech corpus]] for [[speech recognition]] and also [[speaker recognition]]. The dataset contains sound samples of [[Modern Persian]] combination of [[vowel]] and [[consonant]] phonemes from different speakers. Every sound sample contains just one consonant and one vowel So it is somehow labeled in phoneme level. This dataset
Compared to Farsdat speech dataset<ref>Bijankhan, M., Sheikhzadegan, J., Roohani, M. R., Samareh, Y., Lucas, C., & Tebyani, M. (1994). FARSDAT-The Speech Database of Farsi Spoken Language. The Proceedings of the Australian Conference on Speech Science and Technology (Vol. 2, pp. 826–831).</ref> and Persian speech corpus<ref>Halabi, Nawar (2016). Modern Standard Persian Phonetics for Speech Synthesis. University of Southampton, School of Electronics and Computer Science.</ref> it is more easy to use because it is prepared in .mat data files.<ref>{{cite
==Contents==
The corpus is downloadable from its
* .mat data files of sound samples in a 23*6*30000 matrix, in which 23 is number of consonants, 6 is the number of vowels and 30000 is the length of sound sample.
==See also==
Line 13:
==External links==
* [https://
* [https://www.researchgate.net/publication/322298311_Full_Persian_Vowel_recognition_with_MFCC_and_ANN_on_PCVC_speech_dataset PCVC Paper on ResearchGate]
{{Corpus linguistics}}
[[Category:Datasets in machine learning]]
[[Category:Speech recognition]]
[[Category:Speaker recognition]]
|