Table · dataset · 2022
Gigaspeech
Listed in Hugging Face Datasets
Description
Dataset Card for Gigaspeech Dataset Description GigaSpeech is an evolving, multi-domain English speech recognition corpus with 10,000 hours of high quality labeled audio suitable for supervised training. The transcribed audio data is collected from audiobooks, podcasts and YouTube, covering both read and spontaneous speaking styles, and a variety of topics, such as arts, science, sports, etc. Example Usage The training split has several configurations of… See the full description on the dataset page: huggingface.co/datasets/speechcolab/gigaspeech.
Links
Get the data
- Hugging Face dataset page huggingface.co/datasets/speechcolab/gigaspeech ↗
Gated (auto) - request access on the dataset page
landing page · request access · from Hugging Face
Where it is published
- DOI doi.org/10.57967/hf/6261 ↗
DOI / persistent id · from Hugging Face
Documentation and papers
- arXiv:2106.06909 arxiv.org/abs/2106.06909 ↗
publication · from Hugging Face
Catalogue records · 1
- Hub API huggingface.co/api/datasets/speechcolab/gigaspeech ↗
metadata API · from Hugging Face
Topics
- Stated by source
- audio · automatic speech recognition · text · text to audio · text to speech
- From keywords
- Computer Science & AI
- Inferred from text
- Audio 75%
Provenance · 1 source records, 14 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| Hugging Face Datasets | speechcolab/gigaspeech | 9 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · Hugging Face | connector:huggingface@1.0.0 | /gated |
| concepts[field].local:field:computer-science-ai | mapping · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:audio | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:text | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].local:modality:audio | enrichment · Hugging Face | keyword-concept-rules@1.0.0 | title+description (75%) |
| concepts[task].hf_task:automatic-speech-recognition | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:text-to-audio | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:text-to-speech | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| created_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| description | source · Hugging Face | connector:huggingface@1.0.0 | /description |
| license | source · Hugging Face | connector:huggingface@1.0.0 | /tags[license:*] |
| publication_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| title | source · Hugging Face | connector:huggingface@1.0.0 | /id |
| updated_date | source · Hugging Face | connector:huggingface@1.0.0 |