City of Seattle Open Data portal2024 · dataset
Discrimination Case Closures by Month and Type, 2017-PresentA closed discrimination case is an investigation that has been completed at the Seattle Office for Civil Rights (SOCR). This dataset shows closed cases by month and case type.
City of Seattle Open Data portal2021 · dataset
City of Seattle Racial Equity ActionsThis dataset includes the actions the City of Seattle has been taking towards racial equity. We publish it for accountability and transparency. See who to contact, what we'll deliver, and how we plan on meeting our desired outcomes. You can also view this information here: https://www.seattle.gov/rsji/city-racial-equity-actions#/1
Teesside University Research Data Repository2023 · dataset
Handwritten Numbers (HN)Handwritten numbers from one to three digits extracted from public documents. There are 10000 images of different sizes of numbers between 0 and 303 with a imbalanced distribution, there are even numbers without samples. This dataset was created to do proofs with classification models reuse. The images are named with numbers from 0 to 9999, the file labels.csv contain the labels. The number in the
Teesside University Research Data Repository2024 · dataset
BengaliPrintDB: A Repository of Machine-Printed Bengali Documents…Distinguishing between handwritten and machine-printed documents is vital in OCR applications due to varying processing methods.… …Training data selection differs, with diverse datasets for handwritten OCR models and uniform fonts for machine-printed OCR models.…
figshare + Loughborough Research Repository2026 · Astronomical catalogue
LongHisDoc: A Comprehensive Benchmark for Chinese Long Historical Document Understanding…/LongHisDoc_IMG), their OCR results (./OCR_res), and QA pairs (./LongHisDoc.json).</p>…
Teesside University Research Data Repository2026 · dataset
Multilingual AI-Based Dyslexia Detection…This dataset contains 852 character-level handwriting images collected from 100 unique children aged 6–10 years in school settings in Narowal, Punjab, Pakistan, for research on handwriting-based… …Urdu handwriting images were retained without OCR-based correction or linguistic normalization to preserve visually relevant characteristics such as dots, curves, loops, stroke formation…
Borealis + Agri-environmental Research Data Dataverse2022 · dataset · unknown
Fully Annotated Gas Prices of America Dataset for Multi-Metric Extraction in the WildThe fully annotated Gas Prices of America Dataset for Multi-Metric Extraction (GPA4MME) is a large-scale dataset of metric-specific annotations for gas prices and gas signs obtained from Google Streetview throughout the 49 mainland United States of America. The annotations greatly expand the utility of the original GPA dataset for multi-digit, multi-number price detection and context mapping (deno