e-cienciaDatos2026 · dataset · unknown
Minitutorial de OCR con un VLM formato taller: fine-tuning, inferencia y métricas con Gemma 3 4B (ipynb + instrucciones)…El presente recurso se trata de un minitutorial en formato taller que pretende ser una introducción a realizar OCR con un VLM (Vision Language Model) empleando Python, el repositorio… …No se necesitan conocimientos previos, ya que las celdas vienen preparadas para funcionar tras ejecutarse, pero se valoran conocimientos básicos de OCR y Python para sacar el máximo…
IISH Dataverse2026 · dataset · unknown
GLOBALISE Ground Truth for Handwritten Text and Layout RecognitionThis dataset contains Ground Truth PageXML files that were used to finetune the GLOBALISE Handwritten Text Recognition, baseline detection and region detection models (see Related Publications).
IISH Dataverse2026 · dataset · unknown
GLOBALISE Laypa Region Model - August 2023This is a Laypa region detection model that was created to detect and identify regions (such as page number, heading, paragraph, and marginalia) on the scans of the GLOBALISE VOC corpus. It was trained on Ground Truth that is also available in this Dataverse (GLOBALISE Ground Truth for Handwritten Text and Layout Recognition) and applied using the Loghi Handwritten Text Recognition tools to genera
IISH Dataverse2026 · dataset · unknown
GLOBALISE Loghi Handwritten Text Recognition Model – August 2023This is a Loghi Handwritten Text Recognition (HTR) model that was created to transcribe text on the scans of the GLOBALISE VOC corpus. It was finetuned using Ground Truth that is also available in this Dataverse (GLOBALISE Ground Truth for Handwritten Text and Layout Recognition) and applied using the Loghi HTR tools to generate VOC transcriptions v2 - GLOBALISE. The model's hyperparameters are av
e-cienciaDatos2023 · dataset · unknown
UC3M-LP DatasetUC3M-LP is the largest open-source dataset for European license plate detection and recognition and the first one ever for Spanish license plates. It contains 1975 images from 2547 different vehicles and comprising a total of 12757 plate characters.
e-cienciaDatos2023 · dataset · unknown
UC3M-VRI DatasetUC3M-VRI is a dataset for Vehicle Re-Identification based on visual features. It contains 1611 images from 286 different vehicles. Labels include information about make and model of the vehicle.
Borealis + Agri-environmental Research Data Dataverse2022 · dataset · unknown
Fully Annotated Gas Prices of America Dataset for Multi-Metric Extraction in the WildThe fully annotated Gas Prices of America Dataset for Multi-Metric Extraction (GPA4MME) is a large-scale dataset of metric-specific annotations for gas prices and gas signs obtained from Google Streetview throughout the 49 mainland United States of America. The annotations greatly expand the utility of the original GPA dataset for multi-digit, multi-number price detection and context mapping (deno