Data · dataset · 2023
Inv3D: a high-resolution 3D invoice dataset for template-guided single-image document unwarping - Train split part 3 of 4
Listed in RADAR and RADAR4Memory — shown once because both records carry DOI 10.35097/1695
Numerous business workflows involve printed forms, such as invoices or receipts, which are often manually digitalized to persistently search or store the data.
Description
As hardware scanners are costly and inflexible, smartphones are increasingly used for digitalization. Here, processing algorithms need to deal with prevailing environmental factors, such as shadows or crumples.
Current state-of-the-art approaches learn supervised image dewarping models based on pairs of raw images and rectification meshes. The available results show promising predictive accuracies for dewarping, but generated errors still lead to sub-optimal information retrieval. In this paper, we explore the potential of improving dewarping models using additional, structured information in the form of invoice templates.
Read the rest (2 more)
We provide two core contributions: (1) a novel dataset, referred to as Inv3D, comprising synthetic and real-world high-resolution invoice images with structural templates, rectification meshes, and a multiplicity of per-pixel supervision signals and (2) a novel image dewarping algorithm, which extends the state-of-the-art approach GeoTr to leverage structural templates using attention. Our extensive evaluation includes an implementation of DewarpNet and shows that exploiting structured templates can improve the performance for image dewarping.
We report superior performance for the proposed algorithm on our new benchmark for all metrics, including an improved local distortion of 26.1 %. We made our new dataset and all code publicly available at felixhertlein.github.io/inv3d.
Links
Where it is published
- DOI doi.org/10.35097/1695 ↗
DOI / persistent id · from radar service eu de
Catalogue records · 1
- OAI-PMH record radar-service.eu/oai/OAIHandler?verb=GetRecord&metadataPrefix=oai_dc&identifier… ↗
metadata API · from radar service eu de
Topics
- From keywords
- Computer Science & AI · Computer Science & AI · Earth & Environmental Science · Earth & Environmental Science · Optical character recognition · Optical character recognition
- Inferred from text
- Image 75%
Provenance · 2 source records, 12 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| RADAR | 10.35097/1695 | 4 d ago | JSON v1 |
| RADAR4Memory | 10.35097/1695 | 4 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · radar service eu de | connector:radar_service_eu_de@1.0.0 | |
| concepts[field].local:field:computer-science-ai | mapping · radar service eu de | connector:radar_service_eu_de@1.0.0 | |
| concepts[field].local:field:computer-science-ai | mapping · radar4memory radar service eu | connector:radar4memory_radar_service_eu@1.0.0 | |
| concepts[field].local:field:earth-environmental | mapping · radar4memory radar service eu | connector:radar4memory_radar_service_eu@1.0.0 | |
| concepts[field].local:field:earth-environmental | mapping · radar service eu de | connector:radar_service_eu_de@1.0.0 | |
| concepts[method].local:method:ocr | mapping · radar4memory radar service eu | vocabulary-mapper@1.0.0 | keywords['OCR'] |
| concepts[method].local:method:ocr | mapping · radar service eu de | vocabulary-mapper@1.0.0 | keywords['OCR'] |
| concepts[modality].local:modality:image | enrichment · radar service eu de | keyword-concept-rules@1.0.0 | title+description (75%) |
| description | source · radar service eu de | connector:radar_service_eu_de@1.0.0 | /metadata/dc/description |
| license | source · radar service eu de | connector:radar_service_eu_de@1.0.0 | /metadata/dc/rights |
| publication_date | source · radar service eu de | connector:radar_service_eu_de@1.0.0 | |
| title | source · radar service eu de | connector:radar_service_eu_de@1.0.0 | /metadata/dc/title |