Constarium
← Search

Data · dataset · 2026

Estonian Double Modal Corpus sentences

Listed in DATADOI

This dataset contains 739 Estonian double modal constructions extracted from the 2023 Estonian National Corpus.

Description

It was compiled to investigate the distribution and semantics of double modal constructions in Estonian as part of an MA thesis. The dataset aims to include approximately 30 examples of each double modal verb combination where sufficient corpus data were available.

Each record represents a single sentence and includes the corpus token identifier (toknum), the outer (part1) and inner (part2) modal verbs, the ordering of the outer modal, inner modal, and lexical verb (order), the left and right sentence context surrounding the keyword-in-context span (left, kwic, right), a semantic classification (type, with X indicating an unsuitable or exceptional example), and whether the construction is negated (negation, y = negated, n = not negated).

Links

Where it is published

Catalogue records · 1

Topics

Stated by source
Arts and Humanities
From keywords
Humanities
Inferred from text
Linguistics 72%
Provenance · 1 source records, 9 field assertions
SourceKeyLast seenRaw
DATADOIdoi:10.23673/BHLELN8 d agoJSON v1
FieldAssertionExtractorEvidence
concepts[field].anzsrc:group:4704enrichment · datadoi eetaxonomy-embedding@1.1.0title+keywords+description (72%)
concepts[field].dataverse_subject:arts-and-humanitiessource · datadoi eeconnector:datadoi_ee@1.0.0/subjects
concepts[field].local:field:humanitiesmapping · datadoi eeconnector:datadoi_ee@1.0.0/subjects
created_datesource · datadoi eeconnector:datadoi_ee@1.0.0
descriptionsource · datadoi eeconnector:datadoi_ee@1.0.0/description
publication_datesource · datadoi eeconnector:datadoi_ee@1.0.0
titlesource · datadoi eeconnector:datadoi_ee@1.0.0/name
updated_datesource · datadoi eeconnector:datadoi_ee@1.0.0
version_labelsource · datadoi eeconnector:datadoi_ee@1.0.0