Hugging Face Datasets2026 · Image
U(1) Defect Phase-Crossover ExperimentU(1) defect phase-crossover experiment Numerical study of the smallest eigenvalue of the unnormalized scalar connection Laplacian of an n×n open square grid with exactly one phased edge (plus 2D torus and 3D box extensions). Part of a multi-agent project on local frustration vs global spectral visibility. Start here: REPORT.md (v3, post-audit) — results with [T]/[A]/[N]/[C]/[O] evidence labels and
Hugging Face Datasets2026 · Table · CSV
Adopt a BuddyAdopt a Buddy Task: Multiclass ClassificationMissing Values: No
Hugging Face Datasets2026 · Table · Parquet
pigProfessional/so-arm101-stack-green_20260929_160851This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos" ], "shape": [ 6… See the full description on the dataset page: https://huggingface.co/datasets/pigProfessional/so-arm10
Hugging Face Datasets2026 · Table · CSV
Adult Census IncomeAdult Census Income Task: Binary ClassificationMissing Values: Yes
Hugging Face Datasets2026 · Table · CSV
California HousingCalifornia Housing Task: Regression Missing Values: No
Hugging Face Datasets2026 · Table · Parquet
EvalSafe O*NETEvalSafe O*NET 150 documents · 7,500 consensus-labeled questions · 9 candidate models. Snapshot: 2026-09-29. Default reference: consensus. Only questions with an available consensus target and their corresponding documents and final model results are included. The documents are synthetic workplace examples. The reference targets are model-generated, using Astra (gpt-6-astra) and Fable (claude-fabl
Hugging Face Datasets2026 · Table · CSV
Cardiovascular DiseaseCardiovascular Disease Task: Binary ClassificationMissing Values: No
Hugging Face Datasets2026 · Image
datachain/BVD-V-55M-CCBVD-V-55M-CC — the Creative Commons subset Every video in laion/BVD-V-55M-URLs whose YouTube licence is creativeCommon, in the same format as the original. The licence is not part of BVD, so it was resolved for the whole corpus through the YouTube Data API: all 2,119,224 source videos, one videos.list(part=status) call per fifty ids. 15,658 came back creativeCommon, 0.74%. This is a complete censu
Hugging Face Datasets2026 · Text
liuyueyi-8/Elastic-Forcing-training-datasetElastic-Forcing training datasets wan-1.3B-dataset/: 8,682 original videos and paired captions used by experiment 10351. Each dataset folder contains its own videos, caption metadata, training manifest, and provenance. Original training data are kept separate across model scales. wan-14B-dataset/: 5,546 retained videos and paired captions from the 14B step80 training dataset (originally 8,000; 2,4
Hugging Face Datasets2026 · Table · Parquet
Qwen3.8-Max DistillationQwen3.8-Max Distillation A quality-filtered derivative of Qwen3.8-Max Distillation 50K by r0b0tlab, prepared for local training and fine-tuning on consumer hardware. This repository takes the original 49,772-example dataset and produces a substantially smaller training set focused on coding, reasoning, instruction following, and tool use. [!CAUTION] Terms and provenance notice — not cleared for un
Hugging Face Datasets2026 · Table · Parquet
BBuckz/basketball-encyclopediaHugging Face Datasets2026 · Table · Parquet
typesafe/evalsafe-invoice-processingInvoice processing Snapshot: 2026-09-28. 150 cases and 6,874 question instances. Default reference: consensus. Labels are model-generated references. Data Load configuration cases, questions, or run_results; all have a test split. cases: one row per case_id, with the complete input in input_json, descriptive metadata_json, and openai, anthropic, and consensus labelsets. Decisions are grouped by po
Hugging Face Datasets2026 · Text
BaRe-Mem DataBaRe-Mem: Bayesian Reliability Memory for Robust and Adaptive Agent Consultation Overview In multi-agent systems, a central model can consult advisors, but advisor capabilities vary across tasks, and misleading information can make consultation worse than autonomous reasoning. BaRe-Mem is an online Bayesian reliability memory for multi-agent consultation: it estimates each advisor's reliability fr
Hugging Face Datasets2026 · Table · Parquet
dhgottesman/paq-kas-clozeHugging Face Datasets2026 · Table · Parquet
PriceLab Product PricesPriceLab Product Prices I prepared this dataset for PriceLab, my product-price estimation project. It contains 19,991 training examples, 999 validation examples and 1,000 test examples derived from ed-donner/items_lite. I preserved the original split assignments and filtered descriptions containing explicit price expressions before training. What this dataset is for I use product descriptions to t
Hugging Face Datasets2026 · Table · Parquet
haijian06/yellow_cube_brown_box_v1This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos" ], "shape": [ 6… See the full description on the dataset page: https://huggingface.co/datasets/haijian06/yellow_cube_br
Hugging Face Datasets2026 · Table · Parquet
Whittle distillation set (Qwen3.8-27B reasoning traces + top-20 logprobs)Whittle distillation set: Qwen3.8-27B traces with top-20 logprobs About 25,000 complete, quality-filtered answers from Qwen3.8-27B (thinking on), each with the teacher's top-20 log-probabilities at every answer token. It is built for logit-level knowledge distillation: a student can match the teacher's full per-token distribution, not just its sampled text. We use it to train Whittle-Qwen-3.8-35B-
Hugging Face Datasets2026 · Image
VidScribeVidScribe VidScribe is a diagnostic benchmark for visual text in video generation. It has four tasks: T2V (render text from a prompt), R2V (transfer text identity from a reference image), I2V (keep text intact under motion from a first frame), and V2V (edit localized text in an existing video). Every sample is labeled on 12 factor axes (F1–F12). Release status. This repository hosts the public hal
Hugging Face Datasets2026 · Image
VGDL-fMRI — Reason to PlayVGDL-fMRI: Reason to Play Human video-game learning, fMRI recordings, model gameplay and representations for Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners, accepted at NeurIPS 2026. Research code · Interactive results · Original human dataset TL;DR: Explore the replays on the website, or download the human recordings, model features and processed fMRI
Hugging Face Datasets2026 · Image
荆楚文化文物语义分析样本荆楚文化文物语义分析样本 这是用于审核字段设计和语义抽取质量的样本版本,共 116 条记录、35 个核心字段。 数据集另含 enrichment_trial_5 配置:从主表选取 5 条文物进行检索、图像观察与语义补充,共 40 个字段。主表内容未被覆盖。 湖北省博物馆:83 条 荆州博物馆:33 条 图片位于第 2 字段 image_url 两个古籍书影汇总页已拆分为 18 条单书记录 删除了当前来源完全无法填充的 creation_place 和 collection_number 删除派生检索字段 keywords;删除与 archaeological_site 高度重复的 provenance 原文与语义归纳分离;缺少来源的信息保持为空 所有记录目前均为 待人工复核,尚不是最终 604 条全量版本 数据文件 data/artifacts.csv:Dataset Viewer 使用的
Hugging Face Datasets2026 · Text
HUMMBL 40k Multi-Agent Wicked Problems & Coordination CorpusHUMMBL 40k Multi-Agent Wicked Problems & Coordination Corpus A foundational 40,171-event empirical dataset capturing real-world multi-agent coordination, epistemic problem decomposition, failure mode taxonomies, and strategic intelligence surges generated across the HUMMBL autonomous agent fleet. Dataset Overview The dataset provides structured visibility into how autonomous agents navigate comple
Hugging Face Datasets2026 · Table · Parquet
Aaaay-0610/so101-task-1This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos" ], "shape": [ 6… See the full description on the dataset page: https://huggingface.co/datasets/Aaaay-0610/so101-task-1.
Hugging Face Datasets2026 · Text
nabil420/dataHugging Face Datasets2026 · Table · Parquet
YijiaFan/UMM-Reflection-SFT-DataUMM-Reflection SFT Data The reflection-SFT data of UMM-Reflection (Learning Native Reflection in Unified Models). It trains UMM-Reflection-BAGEL-SFT. Research use only, non-commercial. The rows are derived from datasets with different licenses, some of them non-commercial. Each row records its source dataset and license in source_dataset and source_license, and each row follows the terms of its so
Hugging Face Datasets2026 · Table · Parquet
SecondState FAB — Agent Traces and GradingFAB — Agent Traces and Grading 600 completed agent runs: four models × 50 tasks × three trials. Agents investigate a synthetic company's data room and answer financial due-diligence questions. Each row pairs a full execution trace with the task, final answer, grading criteria, pass/fail verdicts and judge explanations. Tasks 041 and 049 are included. The benchmark dataset contains the shared data