Assamese Sentence Aligned Speech Corpus
SKU: LDCIL-233
800x600 Normal 0 false false false EN-US X-NONE HI MicrosoftInternetExplorer4 /* Style Definitions */ table.MsoNormalTable...
View Dataset →SKU: LDCIL-233
800x600 Normal 0 false false false EN-US X-NONE HI MicrosoftInternetExplorer4 /* Style Definitions */ table.MsoNormalTable...
View Dataset →SKU: LDCIL-234
Dataset Description: 69:10:03 hours | 43.3 GB | 40,240 Audio Segments | 450 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-249
08:32:54 hours | 5.6 GB | 5,039 Audio Segments | 61 Speakers The LDC-IL Dogri Sentence Aligned Speech dataset comprises audio files in wav...
View Dataset →SKU: LDCIL-235
Dataset Description: 72:34:52 hours | 45.9 GB | 42,275 Audio Segments | 473 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-244
Dataset Description: 09:21:08 hours | 5.53 GB | 5,676 Audio Segments | 52 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-245
Dataset Description: 11:17:40 hours | 7.27 GB | 6,166 Audio Segments | 53 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-236
Dataset Description: 107:48:50 hours | 69.4 GB | 65,533 Audio Segments | 600 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-237
Dataset Description: 83:19:42 hours | 53.5 GB | 34,091 Audio Segments | 487 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-238
Dataset Description: 41:54:30 hours | 26 GB | 21,412 Audio Segments | 300 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-239
Dataset Description: 123:29:55 hours | 79.6 GB | 89,269 Audio Segments | 451 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-261
116:34:24 hours | 75.9 GB | 60,819 Audio Segments | 589 speakers The LDC-Manipuri Sentence Aligned Speech dataset comprises audio...
View Dataset →SKU: LDCIL-260
116:34:24 hours | 75.9 GB | 60,819 Audio Segments | 589 speakers The LDC-Manipuri Sentence Aligned Speech dataset comprises audio...
View Dataset →SKU: LDCIL-240
Dataset Description: 41:34:04 hours | 26.7 GB | 23,234 Audio Segments | 302 speakers The annotated speech corpus gives wide...
View Dataset →SKU: LDCIL-241
Dataset Description: 43:04:23 hours | 27.7 GB | 21,481 Audio Segments | 346 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-246
Dataset Description: 69:07:50 hours | 44.5 GB | 43,448 Audio Segments | 450 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-259
52:24:51 hours | 34:8 GB | 31,338 Audio Segments | 449 Speakers The LDC-IL Punjabi Sentence Aligned Speech dataset comprises audio files in...
View Dataset →SKU: LDCIL-242
Dataset Description: 74:57:59 hours | 46.4 GB | 48,572 Audio Segments | 433 speakers The annotated speech corpus gives wide range of...
View Dataset →SKU: LDCIL-255
15:38:53 hours | 10.1 GB | 9,548 Audio Segments | 80 Speakers The LDC-IL Telugu Sentence Aligned Speech dataset comprises audio...
View Dataset →SKU: LDCIL-SP-TE-SA-001
Sentence-aligned Telugu speech corpus suitable for speech technology and alignment research.
View Dataset →SKU: LDCIL-243
Dataset Description: 50:09:56 hours | 32.3 GB | 32,384 Audio Segments | 434 speakers The annotated speech corpus gives wide range of...
View Dataset →