Constarium
← Search

Data · dataset · 2026

Coordination of an adverb with a sentence

Listed in REDU - Unicamp Institutional Research Data Repository

Description

The data reported herein, which form part of the FAPESP-funded project (23/16142-0) entitled "Diagnostics for Adverbs: (Micro)variation and Cartography", constitute a systematic test of the entire universal adverbial hierarchy posited by Cinque (1999), in which each adverb is coordinated with a full sentence, across three varieties of Portuguese: namely, Angolan, Brazilian, and Mozambican. This dataset comprises two types of files: 1.

Three files containing the prompts provided to the AI for the generation of the sentences to be judged by native speakers. These files are identified by the word "Prompts" in the title and by the acronyms for each variety: BP (Brazilian Portuguese), AO (Angolan Portuguese), and MOZ (Mozambican Portuguese). Each file contains the prompt carefully prepared by the project executor, Prof.

Read the rest (3 more)

Aquiles Tescari Neto. The prompts specify the nature of the task to be undertaken by the AI, are theoretically oriented toward the target construction, and stipulate the type of sentences to be generated by the AI. In the FAPESP project within which these data were produced, the AI was used only to generate sentences for Brazilian Portuguese and to adapt the generated sentences to the linguistic and cultural reality of Angola and Mozambique. 2.

Three files containing the sentences already judged by native speakers of AO, BP, and MOZ. Each file includes, in addition to the judged sentences, a page with the project metadata. That is, once the data were generated by the AI, they were carefully reviewed by the project executor, Prof.

Aquiles Tescari Neto, and submitted to acceptability/grammaticality judgments by native speakers of the identified varieties. There is one file per language, each with its respective metadata. They are identified by the acronyms AO, BP and MOZ, preceded by "Corpus".

Links

Where it is published

Catalogue records · 1

Topics

Stated by source
Arts and Humanities
From keywords
Humanities
Inferred from text
Linguistic structures 80%
Provenance · 1 source records, 9 field assertions
SourceKeyLast seenRaw
REDU - Unicamp Institutional Research Data Repositorydoi:10.25824/redu/GAAGUR5 d agoJSON v1
FieldAssertionExtractorEvidence
concepts[field].anzsrc:field:470409enrichment · redu unicamp brtaxonomy-embedding@1.1.0title+keywords+description (80%)
concepts[field].dataverse_subject:arts-and-humanitiessource · redu unicamp brconnector:redu_unicamp_br@1.0.0/subjects
concepts[field].local:field:humanitiesmapping · redu unicamp brconnector:redu_unicamp_br@1.0.0/subjects
created_datesource · redu unicamp brconnector:redu_unicamp_br@1.0.0
descriptionsource · redu unicamp brconnector:redu_unicamp_br@1.0.0/description
publication_datesource · redu unicamp brconnector:redu_unicamp_br@1.0.0
titlesource · redu unicamp brconnector:redu_unicamp_br@1.0.0/name
updated_datesource · redu unicamp brconnector:redu_unicamp_br@1.0.0
version_labelsource · redu unicamp brconnector:redu_unicamp_br@1.0.0