Table · dataset · 2026
PubMed-Ophtha
Listed in Hugging Face Datasets
Description
PubMed-Ophtha PubMed-Ophtha is a hierarchical ophthalmic vision-language dataset built from open-access articles in PubMed Central. Figures are extracted directly from the article PDFs at full resolution and decomposed into panels, panel identifiers, and individual images; figure captions are split into panel-level subcaptions. The result is a panel-centric corpus of 102,023 panels from 35,544 figures across 15,842 articles, in which each panel is paired with its subcaption and… See the full description on the dataset page: huggingface.co/datasets/pubmed-ophtha/PubMed-Ophtha.
Links
Where it is published
- Hugging Face dataset page huggingface.co/datasets/pubmed-ophtha/PubMed-Ophtha ↗
landing page · from Hugging Face
Documentation and papers
- arXiv:2605.02720 arxiv.org/abs/2605.02720 ↗
publication · from Hugging Face
Catalogue records · 1
- Hub API huggingface.co/api/datasets/pubmed-ophtha/PubMed-Ophtha ↗
metadata API · from Hugging Face
Topics
- Stated by source
- image feature extraction · image to text · object detection · tabular · text · zero shot image classification
- From keywords
- Computer Science & AI
- Inferred from text
- Image 65%
Provenance · 1 source records, 15 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| Hugging Face Datasets | pubmed-ophtha/PubMed-Ophtha | 7 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · Hugging Face | connector:huggingface@1.0.0 | /gated |
| concepts[field].local:field:computer-science-ai | mapping · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:tabular | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:text | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].local:modality:image | enrichment · Hugging Face | keyword-concept-rules@1.0.0 | title+description (65%) |
| concepts[task].hf_task:image-feature-extraction | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:image-to-text | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:object-detection | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:zero-shot-image-classification | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| created_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| description | source · Hugging Face | connector:huggingface@1.0.0 | /description |
| license_text | source · Hugging Face | connector:huggingface@1.0.0 | |
| publication_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| title | source · Hugging Face | connector:huggingface@1.0.0 | /id |
| updated_date | source · Hugging Face | connector:huggingface@1.0.0 |