Text · dataset · 2022
image captioning language grounding visual semantic
Listed in Hugging Face Datasets
Description
Update: OCT-2023 Add v2 with recent SoTA model swinV2 classifier for both soft/hard-label visual_caption_cosine_score_v2 with person label (0.2, 0.3 and 0.4) Introduction Modern image captaining relies heavily on extracting knowledge, from images such as objects, to capture the concept of static story in the image. In this paper, we propose a textual visual context dataset for captioning, where the publicly available dataset COCO caption (Lin et al., 2014) has been… See the full description on the dataset page: huggingface.co/datasets/AhmedSSabir/Textual-Image-Caption-Dataset.
Links
Where it is published
- Hugging Face dataset page huggingface.co/datasets/AhmedSSabir/Textual-Image-Caption-Dataset ↗
landing page · from Hugging Face
Documentation and papers
- arXiv:2301.08784 arxiv.org/abs/2301.08784 ↗
publication · from Hugging Face
- arXiv:1408.5882 arxiv.org/abs/1408.5882 ↗
publication · from Hugging Face
Catalogue records · 1
- Hub API huggingface.co/api/datasets/AhmedSSabir/Textual-Image-Caption-Dataset ↗
metadata API · from Hugging Face
Topics
- Stated by source
- image classification · image to text · sentence similarity · text · visual question answering
- From keywords
- Computer Science & AI
- Inferred from text
- Image 75%
Provenance · 1 source records, 13 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| Hugging Face Datasets | AhmedSSabir/Textual-Image-Caption-Dataset | 12 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · Hugging Face | connector:huggingface@1.0.0 | /gated |
| concepts[field].local:field:computer-science-ai | mapping · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:text | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].local:modality:image | enrichment · Hugging Face | keyword-concept-rules@1.0.0 | title+description (75%) |
| concepts[task].hf_task:image-classification | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:image-to-text | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:sentence-similarity | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:visual-question-answering | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| created_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| description | source · Hugging Face | connector:huggingface@1.0.0 | /description |
| publication_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| title | source · Hugging Face | connector:huggingface@1.0.0 | /id |
| updated_date | source · Hugging Face | connector:huggingface@1.0.0 |