Table · dataset · 2024
linagora/linto-dataset-audio-ar-tn-augmented
Listed in Hugging Face Datasets
LinTO DataSet Audio for Arabic Tunisian Augmented A collection of Tunisian dialect audio and its annotations for STT task This is the augmented datasets used to train the Linto Tunisian dialect with code-switching STT linagora/linto-asr-ar-tn.
Description
Dataset
Summary
Read the rest (3 more)
Dataset composition Sources Content Types Languages and Dialects Example use (python) License Citations Dataset
Summary
The LinTO DataSet Audio for Arabic Tunisian Augmented is a dataset that builds on LinTO… See the full description on the dataset page: huggingface.co/datasets/linagora/linto-dataset-audio-ar-tn-augmented.
Links
Where it is published
- Hugging Face dataset page huggingface.co/datasets/linagora/linto-dataset-audio-ar-tn-augmented ↗
landing page · from Hugging Face
Documentation and papers
- arXiv:2504.02604 arxiv.org/abs/2504.02604 ↗
publication · from Hugging Face
- arXiv:2309.11327 arxiv.org/abs/2309.11327 ↗
publication · from Hugging Face
Catalogue records · 1
- Hub API huggingface.co/api/datasets/linagora/linto-dataset-audio-ar-tn-augmented ↗
metadata API · from Hugging Face
Topics
- Stated by source
- audio · automatic speech recognition · text · text to audio · text to speech
- From keywords
- Computer Science & AI
- Inferred from text
- Artificial intelligence 70% · Audio 75%
Provenance · 1 source records, 15 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| Hugging Face Datasets | linagora/linto-dataset-audio-ar-tn-augmented | 9 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · Hugging Face | connector:huggingface@1.0.0 | /gated |
| concepts[field].anzsrc:group:4602 | enrichment · Hugging Face | taxonomy-embedding@1.1.0 | title+keywords+description (70%) |
| concepts[field].local:field:computer-science-ai | mapping · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:audio | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:text | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].local:modality:audio | enrichment · Hugging Face | keyword-concept-rules@1.0.0 | title+description (75%) |
| concepts[task].hf_task:automatic-speech-recognition | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:text-to-audio | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:text-to-speech | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| created_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| description | source · Hugging Face | connector:huggingface@1.0.0 | /description |
| license | source · Hugging Face | connector:huggingface@1.0.0 | /tags[license:*] |
| publication_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| title | source · Hugging Face | connector:huggingface@1.0.0 | /id |
| updated_date | source · Hugging Face | connector:huggingface@1.0.0 |