Constarium
← Search

Imaging · dataset · 2025

HumanEval-V/HumanEval-V-Benchmark

Listed in Hugging Face Datasets

Description

HumanEval-V: Benchmarking High-Level Visual Reasoning with Complex Diagrams in Coding Tasks 📄 Paper • 🏠 Home Page • 💻 GitHub Repository • 🏆 Leaderboard • 🤗 Dataset Viewer HumanEval-V is a novel benchmark designed to evaluate the diagram understanding and reasoning capabilities of Large Multimodal Models (LMMs) in programming contexts. Unlike existing benchmarks, HumanEval-V focuses on coding tasks that require sophisticated visual reasoning over… See the full description on the dataset page: huggingface.co/datasets/HumanEval-V/HumanEval-V-Benchmark.

Links

Documentation and papers

Catalogue records · 1

Topics

Stated by source
image · image text to text · text
Inferred from text
Artificial intelligence 70%
Provenance · 1 source records, 12 field assertions
SourceKeyLast seenRaw
Hugging Face DatasetsHumanEval-V/HumanEval-V-Benchmark10 d agoJSON v1
FieldAssertionExtractorEvidence
access_levelsource · Hugging Faceconnector:huggingface@1.0.0/gated
concepts[field].anzsrc:group:4602enrichment · Hugging Facetaxonomy-embedding@1.1.0title+keywords+description (70%)
concepts[field].local:field:computer-science-aimapping · Hugging Faceconnector:huggingface@1.0.0
concepts[modality].hf_modality:imagesource · Hugging Faceconnector:huggingface@1.0.0
concepts[modality].hf_modality:textsource · Hugging Faceconnector:huggingface@1.0.0
concepts[task].hf_task:image-text-to-textsource · Hugging Faceconnector:huggingface@1.0.0/tags[task_categories:*]
created_datesource · Hugging Faceconnector:huggingface@1.0.0
descriptionsource · Hugging Faceconnector:huggingface@1.0.0/description
licensesource · Hugging Faceconnector:huggingface@1.0.0/tags[license:*]
publication_datesource · Hugging Faceconnector:huggingface@1.0.0
titlesource · Hugging Faceconnector:huggingface@1.0.0/id
updated_datesource · Hugging Faceconnector:huggingface@1.0.0