Data Β· dataset Β· 2026
Moncyan/CLVG-Bench
Listed in Hugging Face Datasets
Description
How Far Are Video Models from True Multimodal Reasoning? π CLVG-Bench Directory Structure Here is the directory structure for CLVG-Bench along with descriptions for each folder and file type: CLVG-Bench/ ββ metadata.parquet ββ Element_Editing/ β ββ Background_Modification/ β β ββ 1/ β β ββ 2/ β β ββ ... β ββ Camera_Motion_Editing/ β β ββ 1/ β β ββ 2/ β β ββ ... β ββ Dialogue_Editing/ β β ββ 1/ β β ββ ... β βββ¦ See the full description on the dataset page: huggingface.co/datasets/Moncyan/CLVG-Bench.
Links
Where it is published
- Hugging Face dataset page huggingface.co/datasets/Moncyan/CLVG-Bench β
landing page Β· from Hugging Face
Documentation and papers
- arXiv:2604.19193 arxiv.org/abs/2604.19193 β
publication Β· from Hugging Face
Catalogue records Β· 1
- Hub API huggingface.co/api/datasets/Moncyan/CLVG-Bench β
metadata API Β· from Hugging Face
Topics
- From keywords
- Computer Science & AI
- Inferred from text
- Video 75%
Provenance Β· 1 source records, 8 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| Hugging Face Datasets | Moncyan/CLVG-Bench | 11 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source Β· Hugging Face | connector:huggingface@1.0.0 | /gated |
| concepts[field].local:field:computer-science-ai | mapping Β· Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].local:modality:video | enrichment Β· Hugging Face | keyword-concept-rules@1.0.0 | title+description (75%) |
| created_date | source Β· Hugging Face | connector:huggingface@1.0.0 | |
| description | source Β· Hugging Face | connector:huggingface@1.0.0 | /description |
| publication_date | source Β· Hugging Face | connector:huggingface@1.0.0 | |
| title | source Β· Hugging Face | connector:huggingface@1.0.0 | /id |
| updated_date | source Β· Hugging Face | connector:huggingface@1.0.0 |