Table · dataset · 2022
bigscience-data/roots_indic-ta_pib
Listed in Hugging Face Datasets
Description
ROOTS Subset: roots_indic-ta_pib pib Dataset uid: pib Description Sentence aligned parallel corpus between 11 Indian Languages, crawled and extracted from the press information bureau website. Homepage huggingface.co/datasets/pib preon.iiit.ac.in/~jerin/bhasha/ Licensing Creative Commons Attribution-ShareAlike 4.0 International Speaker Locations Sizes 0.0609 % of total 0.6301 % of indic-hi 3.2610 % of… See the full description on the dataset page: huggingface.co/datasets/bigscience-data/roots_indic-ta_pib.
Links
Get the data
- Hugging Face dataset page huggingface.co/datasets/bigscience-data/roots_indic-ta_pib ↗
Gated (auto) - request access on the dataset page
landing page · request access · from Hugging Face
Catalogue records · 1
- Hub API huggingface.co/api/datasets/bigscience-data/roots_indic-ta_pib ↗
metadata API · from Hugging Face
Topics
- Stated by source
- text
- From keywords
- Computer Science & AI
Provenance · 1 source records, 9 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| Hugging Face Datasets | bigscience-data/roots_indic-ta_pib | 12 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · Hugging Face | connector:huggingface@1.0.0 | /gated |
| concepts[field].local:field:computer-science-ai | mapping · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:text | source · Hugging Face | connector:huggingface@1.0.0 | |
| created_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| description | source · Hugging Face | connector:huggingface@1.0.0 | /description |
| license | source · Hugging Face | connector:huggingface@1.0.0 | /tags[license:*] |
| publication_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| title | source · Hugging Face | connector:huggingface@1.0.0 | /id |
| updated_date | source · Hugging Face | connector:huggingface@1.0.0 |