我的资源
共 246 个数据集
LibriVoxDeEn
Speech Recognition
LibriVoxDeEn

Introduced by Beilharz et al. inLibriVoxDeEn: A Corpus for German-to-English Speech Translation and German Speech Recognition

0 下载 · 0 赞获取 →
Libri-Adapt
Speech RecognitionDomain Adaptation
Libri-Adapt

Introduced by Mathur et al. in Libri-Adapt: A New Speech Dataset for Unsupervised Domain Adaptation

0 下载 · 0 赞获取 →
LaboroTVSpeech
Speech Recognition
LaboroTVSpeech

Introduced by Ando et al. inConstruction of a Large-scale Japanese ASR Corpus on TV Recordings

0 下载 · 0 赞获取 →
Flickr Audio Caption Corpus
Audio signals
Flickr Audio Caption Corpus

The Flickr 8k Audio Caption Corpus contains 40,000 spoken captions of 8,000 natural images.

0 下载 · 1 赞获取 →
EasyCom
speech recognition
EasyCom

Introduced by Donley et al. inEasyCom: An Augmented Reality Dataset to Support Algorithms for Easy Communication in Noisy Environments

0 下载 · 0 赞获取 →
Earnings-21
Named Entity Recognition
Earnings-21

Introduced by Rio et al. inEarnings-21: A Practical Benchmark for ASR in the Wild

0 下载 · 0 赞获取 →
BSTC
speech recognition
BSTC

Introduced by Zhang et al. inBSTC: A Large-Scale Chinese-English Speech Translation Dataset

0 下载 · 0 赞获取 →
CSRC
Speech Recognition
CSRC

Introduced by Yu et al. inThe SLT 2021 children speech recognition challenge: Open datasets, rules and baselines

0 下载 · 0 赞获取 →
ASR-GLUE
Speech Recognition
ASR-GLUE

Introduced by Feng et al. inASR-GLUE: A New Multi-task Benchmark for ASR-Robust Natural Language Understanding

0 下载 · 0 赞获取 →
VoxClamantis
NLP
VoxClamantis

Introduced by Salesky et al. inA Corpus for Large-Scale Phonetic Typology

0 下载 · 0 赞获取 →
RAVDESS
Emotion RecognitionAudio Classification
RAVDESS

Ryerson Audio-Visual Database of Emotional Speech and Song

0 下载 · 0 赞获取 →
NISP
speech recognitionaudio signals
NISP

Introduced by Kalluri et al. inNISP: A Multi-lingual Multi-accent Dataset for Speaker Profiling

0 下载 · 0 赞获取 →