Imaging · dataset · 2024
Vaani
Listed in Hugging Face Datasets
VAANI is an India-representative multi-modal multi-lingual dataset.
Description
The current version (phase 1- 80 districts, phase 2- 85 districts) contains ~31278 hours of spontaenous,image-prompted speech by 156K speakers across 165 districts, talking about 288K images covering 105 languages. From this audio data, 2,122 hours of transcribed data(text) is available, spanning almost evenly across the 165 districts.
Project
Read the rest (1 more)
Vaani, by IISc, Bangalore and ARTPARK, is capturing the true diversity of India’s… See the full description on the dataset page: huggingface.co/datasets/ARTPARK-IISc/Vaani.
Links
Get the data
- Hugging Face dataset page huggingface.co/datasets/ARTPARK-IISc/Vaani ↗
Gated (auto) - request access on the dataset page
landing page · request access · from Hugging Face
Documentation and papers
- arXiv:2603.28714 arxiv.org/abs/2603.28714 ↗
publication · from Hugging Face
Catalogue records · 1
- Hub API huggingface.co/api/datasets/ARTPARK-IISc/Vaani ↗
metadata API · from Hugging Face
Topics
- Stated by source
- audio · automatic speech recognition · image · image to text · text · text to image · text to speech
- From keywords
- Computer Science & AI
Provenance · 1 source records, 18 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| Hugging Face Datasets | ARTPARK-IISc/Vaani | 11 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · Hugging Face | connector:huggingface@1.0.0 | /gated |
| concepts[field].local:field:computer-science-ai | mapping · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:audio | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:image | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:text | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].local:modality:audio | enrichment · Hugging Face | keyword-concept-rules@1.0.0 | title+description (75%) |
| concepts[modality].local:modality:image | enrichment · Hugging Face | keyword-concept-rules@1.0.0 | title+description (65%) |
| concepts[modality].local:modality:text | enrichment · Hugging Face | keyword-concept-rules@1.0.0 | title+description (75%) |
| concepts[task].hf_task:automatic-speech-recognition | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:image-to-text | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:text-to-image | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:text-to-speech | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| created_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| description | source · Hugging Face | connector:huggingface@1.0.0 | /description |
| license | source · Hugging Face | connector:huggingface@1.0.0 | /tags[license:*] |
| publication_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| title | source · Hugging Face | connector:huggingface@1.0.0 | /id |
| updated_date | source · Hugging Face | connector:huggingface@1.0.0 |