{
"version": "1.0",
"query": {
"text": null,
"operator": "and",
"filters": [
{
"field": "concept",
"op": "descendant_of",
"value": "hf_task:text-classification",
"evidence": null
}
],
"evidence_policy": "standard"
},
"sort": null,
"page": {
"size": 25
}
}Hugging Face Datasets2026 · Table · Parquet
Typed Decisions (Japanese)Typed Decisions — 日本語版(typed-decisions-ja v3) LocalLLaMA/typed-decisions の日本語版です。1 つの state(業務の状況)に対して型付きの質問をまとめて答え、それぞれの答えを確率分布で返す課題です。 数値は、すべて実際に測った値です。測っていないものは、そう書いています。 概要 元データ: LocalLLaMA/typed-decisions、revision f7a2487edd7a043a5441a5e9ccc7fe5ddbd9ebe8(Apache-2.0)。 翻訳器: Qwen3.5-35B-A3B(Ollama の qwen3.5:35b-a3b-q4_K_M、digest 3460ffeede5453ead027dbd2f821b12ad0aa3de54630971993babdb2165221f7、Ap
Hugging Face Datasets2026 · Table · Parquet
claimcheck-bench: Agent Success-Claim Verificationclaimcheck-bench A synthetic benchmark for checking whether an AI agent's success claim is supported by its tool results and environment state. An agent can say “done” after a failed write, an action on the wrong record, or an operation that never persisted. This dataset contains 300 labelled agent traces for evaluating detectors that distinguish successful completion from false success claims acr
Hugging Face Datasets2026 · Table · CSV
Haitian Creole Word FrequencyHaitian Creole Word Frequency A word frequency list for Haitian Creole, built from the CMU Haitian newswire corpus. 17,948 unique lowercase words with occurrence counts, sorted by count descending. Provenance This dataset is derived from the Haitian Creole text data released by the Language Technologies Institute at Carnegie Mellon University, Copyright (c) 2010, All Rights Reserved. The CMU LTI H
Hugging Face Datasets2026 · Table · Parquet
EinzzCookie/TikTok-Video-User-DataTikTok Reposts & Authors Dataset This dataset contains archived TikTok reposts and user/author metadata scraped continuously via public TikTok API endpoints. It is split into two Parquet files for easy relational querying. 📄 Dataset Structure The dataset consists of two Parquet files: 1. videos.parquet Contains metadata for individual reposted TikTok videos. Column Type Description id String Uniqu
Hugging Face Datasets2026 · Table · Parquet
ARGUSAt a glance · Quick start · Fields · Reading the labels · Citation ARGUS: Evidence-Grounded Auditing of Identification Assumptions Data release for "Evidence-Grounded Auditing of Identification Assumptions in Climate-Policy Causal Evaluations" (Zhang, Xie, Parra & Correia), to appear at ClimateNLP 2026 (EMNLP 2026 workshop). ARGUS is a structured language-model pipeline that audits the evidence a
Hugging Face Datasets2026 · Text
PublicHearingLDSExtendedPublicHearingLDSExtended Dataset PublicHearingLDSExtended é um dataset estruturado para verificação de alegações (claim verification) e recuperação de evidências em audiências públicas da Câmara dos Deputados do Brasil. Ele estende o dataset original PublicHearingBR_LDS, preservando 100% dos seus documentos e metadados, adicionando uma camada estruturada de alegações (claims) por orador com evidên
Hugging Face Datasets2026 · Image
OneJev-DataThe training data of OneJev: 94,707 typed questions about screens, photos, videos and text, each with its answer. This is 95.5% of the rows OneJev was trained on; rows whose sources do not allow redistribution are left out. Use from datasets import load_dataset ds = load_dataset("OmniJev/OneJev-Data", split="train", streaming=True) row = next(iter(ds)) Each row has a state with <image:N> and <vide
Hugging Face Datasets2026 · Text
HUMMBL 40k Multi-Agent Wicked Problems & Coordination CorpusHUMMBL 40k Multi-Agent Wicked Problems & Coordination Corpus A foundational 40,171-event empirical dataset capturing real-world multi-agent coordination, epistemic problem decomposition, failure mode taxonomies, and strategic intelligence surges generated across the HUMMBL autonomous agent fleet. Dataset Overview The dataset provides structured visibility into how autonomous agents navigate comple
Hugging Face Datasets2026 · Text
ZorQelis Decision Suite (ZDS-1)ZorQelis Decision Suite (ZDS-1) One benchmark for decision models. A decision model reads a situation (the state), a question and a fixed list of options, and returns one option with a calibrated confidence. No free text. ZDS-1 measures that job on 22,864 cases in three tracks, and scores every system with the same harness, the same metrics and the same pricing rule. Cases 22,864 in 3 tracks Metri
Hugging Face Datasets2026 · Table · Parquet
VerQen DecisionsVerQen Decisions The decision training data behind VerQen, an open decision model by ZorQelis AI: 1.62 million typed decisions in five configurations, one per training stage. Each row is a situation, one or more questions with fixed options, and a target distribution over those options. Config Train Validation Test What it teaches decisions 1,406,252 16,528 32,430 (calibration) Apply written polic
Hugging Face Datasets2026 · Table · CSV
Discursos del Congreso de los DiputadosDiscursos del Congreso de los Diputados (muestra balanceada) Descripción Este dataset contiene una muestra de 1088 intervenciones en el pleno del Congreso de los Diputados de España, entre 1996 y 2023. Se ha obtenido a partir del corpus ParlLawSpeech mediante un proceso de limpieza y un muestreo balanceado por año y partido político. Se ha creado como parte de la práctica de la asignatura Descubri
Hugging Face Datasets2026 · Table · Parquet
Web Page Quality EDU Strict-Blind 5196Web Page Quality EDU Strict-Blind 5196 This dataset contains 5,196 English web pages sampled from TeraflopAI/web-page-quality-labels at revision aee0ebeff5bf09e9bd6881b9da581ec6ee77db5c, together with a new strict-blind educational-quality judgment. Sampling The sample uses seed 20260927 and is stratified by the source edu_score: Source score Rows 0 1,000 1 1,000 2 1,000 3 1,000 4 1,000 5 196 (all
Hugging Face Datasets2026 · Table · CSV
Old Games Transcript90s Games Transcript A small English-language corpus of narrative text from classic PC games. This dataset was assembled as a compact research corpus for game studies, digital humanities, discourse analysis, narrative analysis, computational stylistics, and computationally assisted close reading. Dataset Config Records Content caesar3 20 Mission briefings + victory messages diablo2_lod 7 Cinematic
Hugging Face Datasets2026 · Table · CSV
Finance Text ClassificationFinance Text Classification Dataset A lightweight English-language dataset designed for text classification tasks in the finance domain.It contains 1,000 synthetic but realistic loan application texts with requested amounts up to €50,000. The dataset is suitable for: Binary text classification (approved vs. rejected) Zero-shot classification experiments Fine-tuning or evaluating language models on
Hugging Face Datasets2026 · Table · Parquet
rishavk77/amc26-entity-resolution-embeddings-testAMC26 — TEST-Split Business Entity Resolution Embeddings (multilingual-e5-base) Precomputed sentence embeddings for the test split of the Amazon ML Challenge 2026 business entity-resolution task. Same pipeline, same model and same channels as the train-split release, so the two are directly comparable. Purpose: blocking / candidate retrieval. For each Source 1 business, narrow the ~10M-record pool
Hugging Face Datasets2026 · Table · Parquet
BEV 150K Decision MixTürkçe çeviri hakkında Dataset şu an sadece test split içermekte Train verisinin çevirisi devam etmektedir Bu veri seti, avbiswas/bev-decision-150K veri setinin (revision f83fe8b97094112fa305bfc793be07c8f8742282) İngilizceden Türkçeye makine çevirisidir. Bu sürümde yer alan split(ler): test (24386 satır). Satır sayısı, sırası, sütunlar ve Parquet şeması kaynakla birebir aynıdır; yalnızca doğal dil
Hugging Face Datasets2026 · Table · CSV
Saeid Homayoun Accounting and Audit AI PortfolioSaeid Homayoun — Accounting & Audit AI Portfolio A curated starting point for public work in AI-enabled accounting, auditing, assurance, finance, SEC/XBRL, ICFR, CAM/KAM, ESG and governance. Start here NAAIL OpenLab SEC 10-Company Accounting Panel Kimi + DeepSeek Accounting/Finance/Audit Stack SEC CAM/ICFR/Governance/ESG Frankenstein Accounting/Audit Benchmark The machine-readable portfolio.csv ra
Hugging Face Datasets2026 · Table · Parquet
Agent Tool Decisions 180KAgent Tool Decisions 180K 180,000 typed decisions that sit behind every agent step. Should the agent call a tool, or answer in text? Which tool? Are the arguments complete? Which tool-using response is better? An agent spends most of its time making these small decisions, and a frontier LLM is an expensive way to make them. They are typed: a question, a fixed set of choices, one correct answer. Th
Hugging Face Datasets2026 · Table · Parquet
Federal Reserve Beige BookFederal Reserve Beige Book Every edition of the Beige Book, the Federal Reserve's Summary of Commentary on Current Economic Conditions by Federal Reserve District, from the first, May 20, 1970, when the Board's pages call it the Redbook, to the latest. One row per section: the national summary, each of the twelve districts, and the one special report (May 18, 1983). In the words of the Board's Bei
Hugging Face Datasets2026 · Text
Decida synthetic typed decisionsDecida synthetic typed decisions 9,879 synthetic training records for typed decisions: a short situation (the state) and one question about it, with the answer given as a probability distribution. It was generated with deepseek-ai/DeepSeek-V4-Flash-0731, and it is one of the two data sources used to train DecidaBERT-large. The three question types match the typed-decisions format: type records wha
Hugging Face Datasets2026 · Table · Parquet
Self-Healing LocatorsSelf-Healing Locators (v0.1) UI tests find elements with locators like #login-submit, [data-testid="save"] or /html/body/main/form/button[1]. When the page changes, those locators break. Self-healing tools try to find the element again automatically, but there has been no public dataset to train or measure them on. This dataset has 6,000 before/after web page pairs. Each one records a target eleme
Hugging Face Datasets2026 · Text
Sev security evidence collectionSev security evidence collection Sev-4B v0.3.0 research update The v0.3.0 response-policy checkpoint improves authored policy decisions from 127/175 to 147/175 across 35 held-out source programs. At its calibration-selected alert threshold, it flags 14/140 permitted decisions, a 10% false-alert rate on this panel. 26/28 registered checks pass. The DNS diagnostic remains 21/32, with one repaired an
Hugging Face Datasets2026 · Text
Screenplay Format Edge CasesScreenplay Format Edge Cases 48 original Fountain specimens in 24 contrast pairs — two near-identical inputs per pair, at the points where the Fountain syntax leaves a choice. In 17 pairs the one difference changes how the lines are classified. In the other 7 it changes the surface and the labels hold: a lowercase scene prefix, a cue extension, a non-Latin cue, escaped characters, a dual-dialogue
Hugging Face Datasets2026 · Text
jev-benchMachine translation (en → tr) of Praveenrajus/jev-bench @ b41e6b2f68a429608e1c0d354324d0bebaec7971 by qwen3.8-flash-next (prompt v1-json-bfb402e8). Included split(s): test (22,773 rows). Labels, ids, soft labels and metadata are unchanged; see manifest.json. jev-bench Real human-labeled data, reformatted into System One questions — with human label distributions wherever they exist. 22 configs · 1
Hugging Face Datasets2026 · Table · Parquet
BEV Decision MixBEV Decision Mix Formerly avbiswas/bev-decision-150K; the old ID redirects here. The default config holds 150,000 English-language decision rows drawn from existing datasets, public records, and text-based game environments. Each row presents source-grounded input state and one or more bounded decisions inspired by the JEV CHOICE, NOUL, and SCORE contracts. The mixture aims to train models that se