Language Resource Search - SHACHI: Language Resource Metadata Database

Language resource #: 3330 Results 781 - 790 of 2023

C-001372: Mandarin Chinese Desktop Speech Recognition Corpus - Stock (70 people)
Desktop/Microphone
This corpus comprises 1,586 entries uttered by 70 speakers of different dialects, ages and various educational levels (38 males and 32 females), recorded through head-mounted noise-canceling microphone. The database comprises 4,199 items. Speech samples are stored as a sequence of 16-bit 22.05kHz WAV for a total of 5.1 hours of speech. The total capacity of the data is 776 Mb.
Each speaker read 60 stocks. Text files are stored in Unicode format. All data have been proofread manually.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.
C-001373: Mandarin Chinese Desktop Speech Recognition Corpus - Stock (849 people)
Desktop/Microphone
This corpus comprises 1,584 entries uttered by 849 speakers of different dialects, ages and various educational levels (420 males and 429 females), recorded over 2 channels (Mic1: SHURE SM58; Mic2: Labtec Axis-002). The database comprises 13,600 stocks per channel. Speech samples are stored as a sequence of 16-bit 44.1kHz WAV for 20 hours of speech per channel. The total capacity of the data is 12 Gb.
Each speaker read 16 items. Text files are stored in Unicode format. All data have been proofread manually.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.
C-001374: Mandarin Chinese Desktop Speech Recognition Corpus - Stock、 Person Name 、Digit String、Simple Chinese sentences、Spontaneous Speech (50 people)
Desktop/Microphone
This corpus comprises 8,206 entries including stocks, person names, digit strings and 8,511 speech files composed of spontaneous speech, uttered by 50 speakers of different dialects, ages and various educational levels (22 males and 28 females), recorded from a stand microphone (SHURE SM58). Speech samples are stored as a sequence of 16-bit 44.1kHz WAV for a total of 24 hours of speech. The total capacity of the data is 7 Gb.
Text files are stored in Unicode format. All data have been proofread manually.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.
C-001375: Mandarin Chinese Telephone Speech Recognition Corpus - Digit String (649 people)
Desktop/Microphone
This corpus comprises 750 entries uttered by 649 speakers of different dialects, ages and various educational levels (340 males and 309 females), recorded over the fixed telephone network. The database comprises 9,750 digit strings. Speech samples are stored as a sequence of 16-bit 8kHz WAV for a total of 16.28 hours of speech.
Each speaker read 15 items. Text files are stored in Unicode format. All data have been proofread manually.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.
C-001376: Mandarin Chinese Telephone Speech Recognition Corpus - Digit String
Desktop/Microphone
This corpus comprises 5,309 entries uttered by 265 speakers of different dialects, ages and various educational levels (134 males and 131 females), recorded over the fixed telephone network. The database comprises 7,606 Chinese digit strings. Speech samples are stored as a sequence of 16-bit 8kHz WAV for a total of 11.8 hours of speech. The total capacity of the data is 648 Mb.
Each speaker read 25-30 items. Text files are stored in Unicode format. All data have been proofread manually.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.
C-001377: Mandarin Chinese Telephone Speech Recognition Corpus - Person Name (649 people)
Desktop/Microphone
This corpus comprises 2,250 entries uttered by 649 speakers of different dialects, ages and various educational levels (340 males and 309 females), recorded over the fixed telephone network. The database comprises 9,750 person names. Speech samples are stored as a sequence of 16-bit 8kHz WAV for a total of 13.97 hours of speech.
Each speaker read 15 items. Text files are stored in Unicode format. All data have been proofread manually.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.
C-001378: Mandarin Chinese Telephone Speech Recognition Corpus - Person Name, Place Name (Mobile telephone 265)
Desktop/Microphone
This corpus comprises 6,952 entries uttered by 265 speakers of different dialects, ages and various educational levels (134 males and 131 females), recorded over the mobile telephone network. The database comprises 13,942 Chinese personal names and place names. Speech samples are stored as a sequence of 16-bit 8kHz WAV for a total of 17.6 hours of speech. The total capacity of the data is 964 Mb.
Each speaker read 15-30 items. Text files are stored in Unicode format. All data have been proofread manually.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.
C-001379: Mandarin Chinese Telephone Speech Recognition Corpus SMS (Mobile telephone 64)
Desktop/Microphone
This corpus comprises 1,079 entries uttered by 64 speakers of different dialects, ages and various educational levels (52 males and 12 females), recorded over the mobile telephone network. The database comprises 3,190 Chinese short messages (SMS). Speech samples are stored as a sequence of 16-bit 8kHz WAV for a total of 3 hours of speech. The total capacity of the data is 161 Mb.
Each speaker read 50 items. Text files are stored in Unicode format. All data have been proofread manually.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.
C-001380: Mandarin Chinese Telephone Speech Recognition Corpus - Simple Chinese sentences (650 people)
Desktop/Microphone
This corpus comprises 8,011 entries uttered by 650 speakers of different dialects, ages and various educational levels (340 males and 310 females), recorded over the fixed telephone network. The database comprises 80,750 simple Chinese sentences. Speech samples are stored as a sequence of 16-bit 8kHz WAV for a total of 134 hours of speech.
400 speakers read 120 items, 250 speakers read 131 items. Text files are stored in Unicode format. All data have been proofread manually.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.
C-001381: Mandarin Chinese Telephone Speech Recognition Corpus - Spontaneous Speech (649 people)
Desktop/Microphone
This corpus comprises spontaneous speech (elicited) from 649 speakers of different dialects, ages and various educational levels (340 males and 309 females), who uttered 40 different topics in a working environment, recorded over the fixed telephone network. Speech samples are stored as a sequence of 16-bit 8kHz WAV for a total of 143.05 hours of speech.
Each speaker read 15 items. Text files are stored in Unicode format. All data have been proofread manually. The total capacity of the data is 7.67 Gb.
The transcriptions include non-speech markers (background noise, background speech, speaker sounds) as well as markers for mispronunciation, channel distortions, words left-out and duplicates.
The corpus aims to be applied to the testing and telephone natural speech recognition system.

SHACHI - Language Resource Metadata Database