Constarium
← Search

Table · dataset · 2024

longvideobench

Listed in Hugging Face Datasets

Dataset Card for LongVideoBench Large multimodal models (LMMs) are handling increasingly longer and more complex inputs.

Description

However, few public benchmarks are available to assess these advancements. To address this, we introduce LongVideoBench, a question-answering benchmark with video-language interleaved inputs up to an hour long.

It comprises 3,763 web-collected videos with subtitles across diverse themes, designed to evaluate LMMs on long-term multimodal understanding. The… See the full description on the dataset page: huggingface.co/datasets/longvideobench/LongVideoBench.

Links

Get the data

Documentation and papers

Catalogue records · 1

Topics

Inferred from text
Machine learning 69%
Provenance · 1 source records, 13 field assertions
SourceKeyLast seenRaw
Hugging Face Datasetslongvideobench/LongVideoBench8 d agoJSON v1
FieldAssertionExtractorEvidence
access_levelsource · Hugging Faceconnector:huggingface@1.0.0/gated
concepts[field].anzsrc:group:4611enrichment · Hugging Facetaxonomy-embedding@1.1.0title+keywords+description (69%)
concepts[field].local:field:computer-science-aimapping · Hugging Faceconnector:huggingface@1.0.0
concepts[modality].hf_modality:tabularsource · Hugging Faceconnector:huggingface@1.0.0
concepts[modality].hf_modality:textsource · Hugging Faceconnector:huggingface@1.0.0
concepts[task].hf_task:multiple-choicesource · Hugging Faceconnector:huggingface@1.0.0/tags[task_categories:*]
concepts[task].hf_task:visual-question-answeringsource · Hugging Faceconnector:huggingface@1.0.0/tags[task_categories:*]
created_datesource · Hugging Faceconnector:huggingface@1.0.0
descriptionsource · Hugging Faceconnector:huggingface@1.0.0/description
licensesource · Hugging Faceconnector:huggingface@1.0.0/tags[license:*]
publication_datesource · Hugging Faceconnector:huggingface@1.0.0
titlesource · Hugging Faceconnector:huggingface@1.0.0/id
updated_datesource · Hugging Faceconnector:huggingface@1.0.0