Hugging Face Datasets2026 · Table · Parquet
OpenJevData-140kDataset Card for OpenJevData-140k Dataset Summary OpenJevData-140k is a curated release of the data collection used to train OpenJev-4B. It contains 146,738 decision-making examples across 19 task categories, organized into SFT and RL splits. Each example presents a state, a question, and a request-specific set of natural-language options. The data include hard answers and soft probability distrib
Hugging Face Datasets2026 · Image
U(1) Defect Phase-Crossover ExperimentU(1) defect phase-crossover experiment Numerical study of the smallest eigenvalue of the unnormalized scalar connection Laplacian of an n×n open square grid with exactly one phased edge (plus 2D torus and 3D box extensions). Part of a multi-agent project on local frustration vs global spectral visibility. Start here: REPORT.md (v3, post-audit) — results with [T]/[A]/[N]/[C]/[O] evidence labels and
Hugging Face Datasets2026 · Table · CSV
Adopt a BuddyAdopt a Buddy Task: Multiclass ClassificationMissing Values: No
Hugging Face Datasets2026 · Table · Parquet
Typed Decisions (Japanese)Typed Decisions — 日本語版(typed-decisions-ja v3) LocalLLaMA/typed-decisions の日本語版です。1 つの state(業務の状況)に対して型付きの質問をまとめて答え、それぞれの答えを確率分布で返す課題です。 数値は、すべて実際に測った値です。測っていないものは、そう書いています。 概要 元データ: LocalLLaMA/typed-decisions、revision f7a2487edd7a043a5441a5e9ccc7fe5ddbd9ebe8(Apache-2.0)。 翻訳器: Qwen3.5-35B-A3B(Ollama の qwen3.5:35b-a3b-q4_K_M、digest 3460ffeede5453ead027dbd2f821b12ad0aa3de54630971993babdb2165221f7、Ap
Hugging Face Datasets2026 · Table · CSV
Adult Census IncomeAdult Census Income Task: Binary ClassificationMissing Values: Yes
Hugging Face Datasets2026 · Table · Parquet
claimcheck-bench: Agent Success-Claim Verificationclaimcheck-bench A synthetic benchmark for checking whether an AI agent's success claim is supported by its tool results and environment state. An agent can say “done” after a failed write, an action on the wrong record, or an operation that never persisted. This dataset contains 300 labelled agent traces for evaluating detectors that distinguish successful completion from false success claims acr
Hugging Face Datasets2026 · Text
JibayAi/jibay_jcf_baseJibay Persian Chat Dataset 📖 Overview Jibay Persian Chat Dataset is a large-scale, conversational instruction dataset primarily written in Persian (Farsi), with a limited amount of English content mixed in. The dataset was curated and formatted specifically to be compatible with the Jibay Chat Format (JCF) — a structured, role-based prompt format designed by JibayAI for training, fine-tuning, and
Hugging Face Datasets2026 · Text
KaliBench-VerifiedDataset Card for KaliBench KaliBench: A Fine-Grained Benchmark for Cybersecurity Tool Use on Kali Linux with Runtime-Free Verifiable Rewards (NeurIPS 2026 Evaluations and Datasets Track) Authors: Pengfei Li1*, Naufal Suryanto1*, Sicheng Zhang1, Muzammal Naseer1,2 1Khalifa University, 2University of Western Australia *Equal contribution 💻 GitHub Code | 📊 Dataset Dataset summa
Hugging Face Datasets2026 · Table · CSV
Haitian Creole Word FrequencyHaitian Creole Word Frequency A word frequency list for Haitian Creole, built from the CMU Haitian newswire corpus. 17,948 unique lowercase words with occurrence counts, sorted by count descending. Provenance This dataset is derived from the Haitian Creole text data released by the Language Technologies Institute at Carnegie Mellon University, Copyright (c) 2010, All Rights Reserved. The CMU LTI H
Hugging Face Datasets2026 · Table · CSV
LIFT-VistaLIFT-Vista LIFT-Vista is a dataset with large camera viewpoint changes and joint camera-layout annotations. Overview camera.csv camlayout.csv Clips 120,898 58,272 (a subset of the camera clips) Annotations camera trajectory, caption camera trajectory, caption, last-frame layout, per-frame box tracks Every clip has 81 frames at 16 fps (about 5 s) at the native resolution of its source video (98% ar
Hugging Face Datasets2026 · Text
Nickyang/RazorCalRazorCal RazorCal is a general-purpose calibration corpus released with RAZOR. RazorCal.json contains 2,048 samples across seven domains (about 8 MB). Each record includes input messages and source attribution. The records carry no architecture-specific fields, so the corpus also suits FFN or layer pruning, perplexity measurement and other calibration-based compression; RAZOR uses it to estimate p
Hugging Face Datasets2026 · Table · Parquet
EvalSafe O*NETEvalSafe O*NET 150 documents · 7,500 consensus-labeled questions · 9 candidate models. Snapshot: 2026-09-29. Default reference: consensus. Only questions with an available consensus target and their corresponding documents and final model results are included. The documents are synthetic workplace examples. The reference targets are model-generated, using Astra (gpt-6-astra) and Fable (claude-fabl
Hugging Face Datasets2026 · Table · CSV
Cardiovascular DiseaseCardiovascular Disease Task: Binary ClassificationMissing Values: No
Hugging Face Datasets2026 · Text
English and Yoruba Cassava SpeechEnglish and Yoruba Cassava Speech Dataset summary This dataset contains 200 short audio clips, totaling about 30.2 minutes: 50 questions and 50 answers in English, and the corresponding 50 questions and 50 answers in Yoruba. The topic is cassava cultivation, especially crop diseases, pests, planting, equipment, and processing. These are prepared prompts and answers recorded by two adults living in
Hugging Face Datasets2026 · Text
GSM8K-UZ (Cyrillic)GSM8K-UZ (Cyrillic) Uzbek GSM8K in the Cyrillic script. Part of a parallel pair: kurbanovxurshidbek/gsm8k-uz-lat and kurbanovxurshidbek/gsm8k-uz-cyr. Both contain exactly the same problems (same idx), differing only in script. Split Rows train 7417 test 1308 Source and construction Latin text: NeuronUz/gsm8k-uz, a machine translation of openai/gsm8k (main) into Uzbek Latin. Cyrillic text: automati
Hugging Face Datasets2026 · Table · Parquet
EinzzCookie/TikTok-Video-User-DataTikTok Reposts & Authors Dataset This dataset contains archived TikTok reposts and user/author metadata scraped continuously via public TikTok API endpoints. It is split into two Parquet files for easy relational querying. 📄 Dataset Structure The dataset consists of two Parquet files: 1. videos.parquet Contains metadata for individual reposted TikTok videos. Column Type Description id String Uniqu
Hugging Face Datasets2026 · Table · Parquet
ARGUSAt a glance · Quick start · Fields · Reading the labels · Citation ARGUS: Evidence-Grounded Auditing of Identification Assumptions Data release for "Evidence-Grounded Auditing of Identification Assumptions in Climate-Policy Causal Evaluations" (Zhang, Xie, Parra & Correia), to appear at ClimateNLP 2026 (EMNLP 2026 workshop). ARGUS is a structured language-model pipeline that audits the evidence a
Hugging Face Datasets2026 · Text
HABIT-BenchHABIT-Bench Evaluating Habit Induction from Longitudinal Weak Evidence in Agent Memory HABIT-Bench evaluates whether a memory system can infer a latent habit from repeated, individually incomplete observations and apply it only when the current context supports it. Tasks test weak-evidence induction, applicability boundaries, local exceptions, temporal changes, and the distinction between user-end
Hugging Face Datasets2026 · Text
PublicHearingLDSExtendedPublicHearingLDSExtended Dataset PublicHearingLDSExtended é um dataset estruturado para verificação de alegações (claim verification) e recuperação de evidências em audiências públicas da Câmara dos Deputados do Brasil. Ele estende o dataset original PublicHearingBR_LDS, preservando 100% dos seus documentos e metadados, adicionando uma camada estruturada de alegações (claims) por orador com evidên
Hugging Face Datasets2026 · Table · Parquet · gated
krsKRS scrape every entry in Poland's National Court Register (Krajowy Rejestr Sądowy), about 1.2 million companies and associations, parsed from the full extract PDFs and updated weekly. from datasets import load_dataset krs = load_dataset("tiagozip/krs", split="train") format one parquet file, data/krs.parquet, one row per KRS number. column type notes krs string 10 digit KRS number register string
Hugging Face Datasets2026 · Table · Parquet
protodotdesign/magicbox-v1MagicBox v1 693,376 records for extraction, choice, binary classification (noul), and ordinal scoring. Format: minifield.magicbox/1.0. request_json contains the source text and field definitions. targets_json contains supervision. Parse these columns with json.loads. Source attribution, revisions, licenses, and transformations are recorded in sources.json. Dataset counts are recorded in report.jso
Hugging Face Datasets2026 · Text
GSM8K-UZ (Latin)GSM8K-UZ (Latin) Uzbek GSM8K in the Latin script. Part of a parallel pair: kurbanovxurshidbek/gsm8k-uz-lat and kurbanovxurshidbek/gsm8k-uz-cyr. Both contain exactly the same problems (same idx), differing only in script. Split Rows train 7417 test 1308 Source and construction Latin text: NeuronUz/gsm8k-uz, a machine translation of openai/gsm8k (main) into Uzbek Latin. This dataset reproduces the L
Hugging Face Datasets2026 · Text
Agentic Coding Chain-of-Thought Dataset🤖 Agentic Coding CoT Dataset A high-quality supervised fine-tuning (SFT) dataset for training agentic coding assistants with Chain-of-Thought reasoning capabilities. 📋 Dataset Description This dataset was created by processing and distilling ~20GB of GitHub crawl data using Minimax-M2 to generate structured, reasoning-rich coding examples. Each sample demonstrates systematic problem-solving with e
Hugging Face Datasets2026 · Image
datachain/BVD-V-55M-CCBVD-V-55M-CC — the Creative Commons subset Every video in laion/BVD-V-55M-URLs whose YouTube licence is creativeCommon, in the same format as the original. The licence is not part of BVD, so it was resolved for the whole corpus through the YouTube Data API: all 2,119,224 source videos, one videos.list(part=status) call per fifty ids. 15,658 came back creativeCommon, 0.74%. This is a complete censu
Hugging Face Datasets2026 · Text
lucasfrag/fact-checking-abstention-saes-dataFact-Checking Abstention SAEs — data Companion data for lucasfrag/fact-checking-abstention-saes. Per base model (e.g. llama-3.1-8b-instruct/) file content examples.jsonl the exact ordered list of training claims (VitaminC + FEVER train splits, gold evidence), one JSON per line: claim, evidence (list), label (SUP/REF/NEI), source. Line i is prompt i of the harvest, so activations can be regenerated