我的资源
共 246 个数据集
GigaSpeech
Speech Recognition
GigaSpeech

An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

0 下载 · 0 赞获取 →
KeSpeech
Speech RecognitionSpeaker Identification
KeSpeech

An Open Source Speech Dataset of Mandarin and Its Eight Subdialects

0 下载 · 0 赞获取 →
The Spoken Wikipedia Corpora
spoken language regconition
The Spoken Wikipedia Corpora

The SWC is a corpus of aligned Spoken Wikipedia articles from the English, German, and Dutch Wikipedia.

0 下载 · 0 赞获取 →
Parkinson Speech Dataset
Parkinson Speech Dataset

Parkinson Speech Dataset is an audio dataset consisting of recordings of 20 Parkinson's Disease (PD) patients and 20 healthy subjects.

0 下载 · 0 赞获取 →
PCVC
Audio Signals
PCVC

Persian Consonant Vowel Combination

0 下载 · 0 赞获取 →
FluencyBank
Speech Recognition
FluencyBank

FluencyBank is a shared database for the study of fluency development.

0 下载 · 0 赞获取 →
ST-CMDS
ST-CMDS

A free Chinese Mandarin corpus by Surfingtech (www.surfing.ai), containing utterances from 855 speakers, 102600 utterances

0 下载 · 0 赞获取 →
TAL_CSASR中英文混合语音数据集
Speech Recognition
TAL_CSASR中英文混合语音数据集

该数据集为好未来英语课授课音频,包含中英文混合讲话的情况,每条音频只有一位说话人。(文件63.36G)

0 下载 · 0 赞获取 →
TAL_SER语音情感数据集
Emotion recognition
TAL_SER语音情感数据集

语音情感数据集为好未来老师上课音频,共包含4541条音频,总时长12.5小时。

0 下载 · 0 赞获取 →
TAL_ASR语音识别数据集
TAL_ASR语音识别数据集

语音识别数据集为好未来线上课程的老师授课音频,涵盖语文、数学两门学科。

0 下载 · 0 赞获取 →
Deeply vocal characterizer
audio signals
Deeply vocal characterizer

Deeply vocal characterizer is a human nonverbal vocalization dataset.

0 下载 · 0 赞获取 →
Parent-Child vocal interaction
Speech Recognition
Parent-Child vocal interaction

Deeply Parent-Child Vocal Interaction contains the interaction of 24 pairs of parent and child(total 48 speakers), such as reading fairy tales, singing children’s songs, conversing, and others, is recorded.

0 下载 · 0 赞获取 →