Constarium
← Search

Data · dataset · 2024

Preprocessing the English, German, Spanish, and Portuguese Corpora of the European Literary Text Collection (ELTeC) Using the AVOBMAT Multilingual Text Mining Tool

Listed in CKAN @ IoT Lab

The results of the preprocessing the English, German, Spanish and Portuguese corpora of the European Literary Text Collection (ELTeC) with the AVOBMAT (Analysis and Visualization of Bibliographic Metadata and Texts) multilingual research tool.

Description

With AVOBMAT, this data can be used to perform a wide range of dynamic text and data mining tasks. For more information, see avobmat.hu .

Links

Topics

Inferred from text
Text 75%
Provenance · 1 source records, 10 field assertions
SourceKeyLast seenRaw
CKAN @ IoT Lab687f5496-ba80-49d0-b77a-27c4124b17a35 d agoJSON v1
FieldAssertionExtractorEvidence
concepts[field].anzsrc:field:460208mapping · ckan iotlab comvocabulary-mapper@1.0.0keywords['natural language processing']
concepts[field].local:field:computer-science-aimapping · ckan iotlab comconnector:ckan_iotlab_com@1.0.0
concepts[field].local:field:engineeringmapping · ckan iotlab comconnector:ckan_iotlab_com@1.0.0
concepts[field].local:field:humanitiesmapping · ckan iotlab comconnector:ckan_iotlab_com@1.0.0
concepts[modality].local:modality:textenrichment · ckan iotlab comkeyword-concept-rules@1.0.0title+description (75%)
created_datesource · ckan iotlab comconnector:ckan_iotlab_com@1.0.0
descriptionsource · ckan iotlab comconnector:ckan_iotlab_com@1.0.0/notes
publication_datesource · ckan iotlab comconnector:ckan_iotlab_com@1.0.0
titlesource · ckan iotlab comconnector:ckan_iotlab_com@1.0.0/title
updated_datesource · ckan iotlab comconnector:ckan_iotlab_com@1.0.0