Data · dataset · 2026
Package for testing coherence among listeners tasked with identifying prosodic functions and obtaining cross-correlations between prosodic functions and parameters
Listed in REDU - Unicamp Institutional Research Data Repository
Description
This Dataset contains the materials needed for conducing coherence analyses of six groups of five listeners tasked with identifying prosodic functions and cross-correlation analyses dedicated to finding consistent cross-correlations between the occurrence of stressed V-V units of prosodic functions and values for prosodic parameters. The materials consist in two R scripts, one dedicated to the coherence analyses, the other to the cross-correlation analyses, along with text files containing the data needed for running the analyses and a README file which offers some helpful information on how to use this Dataset.
The data for the coherence analyses were obtained through the perceptual testing of lay listeners tasked with listening to audio excerpts and annotating the words listened as realizations of focus, terminal boundaries and non-terminal boundaries. Each of the six groups of twelve to thirteen excerpts were listened to and annotated by the same group of five listeners. The aim was to identify those words consistently listened as realizations of these prosodic functions by lay listeners.
Read the rest (3 more)
The coherence analyses were conduced to verify if these listeners showed agreement greater than chance in their annotations, that is, if the similarities in their annotations are due to the identification of prosodic functions and not to chance. For these analyses, the annotations of each group of listeners were converted into binary tables where lines refer to words in the excerpts and columns to each of the group's five listeners; in these tables, a 1 indicates that a word was identified as a prosodic function by that listener, and a 0 indicates that it was not.
One table was generated for each of the three prosodic functions the listeners were tasked with identifying. These tables are the .txt files in the .zip "coherence_analyses" file. The data for the cross-correlation analyses were obtained through the extraction of values for prosodic-acoustic parameters from the audio files and through the agreement between listeners marking prosodic functions.
The aim of these analyses was to identify the prosodic characteristic of words consistently listened as realizing prosodic functions. The .txt file in the .zip "crosscorrelation_analyses" file is the resulting table, which contains values for prosodic-acoustic parameters for each V-V unit (the unit comprised between the onset of a vowel and the onset of the following vowel) in the audio corpus, as well as the indication if that V-V unit is the stressed V-V unit of a word annotated as realizing a prosodic function by at least four out of five listeners.
Links
Where it is published
- Dataverse dataset page redu.unicamp.br/dataset.xhtml?persistentId=doi%3A10.25824%2Fredu%2FSRMJP8 ↗
landing page · from redu unicamp br
- DOI doi.org/10.25824/redu/srmjp8 ↗
DOI / persistent id · from redu unicamp br
Catalogue records · 1
- Dataverse API redu.unicamp.br/api/datasets/:persistentId/?persistentId=doi%3A10.25824%2Fredu… ↗
metadata API · from redu unicamp br
Topics
- Stated by source
- Arts and Humanities · Other
- From keywords
- Humanities
- Inferred from text
- Audio 75% · Linguistics 74% · Tabular 65% · Text 75%
Provenance · 1 source records, 13 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| REDU - Unicamp Institutional Research Data Repository | doi:10.25824/redu/SRMJP8 | 5 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| concepts[field].anzsrc:group:4704 | enrichment · redu unicamp br | taxonomy-embedding@1.1.0 | title+keywords+description (74%) |
| concepts[field].dataverse_subject:arts-and-humanities | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /subjects |
| concepts[field].dataverse_subject:other | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /subjects |
| concepts[field].local:field:humanities | mapping · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /subjects |
| concepts[modality].local:modality:audio | enrichment · redu unicamp br | keyword-concept-rules@1.0.0 | title+description (75%) |
| concepts[modality].local:modality:tabular | enrichment · redu unicamp br | keyword-concept-rules@1.0.0 | title+description (65%) |
| concepts[modality].local:modality:text | enrichment · redu unicamp br | keyword-concept-rules@1.0.0 | title+description (75%) |
| created_date | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | |
| description | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /description |
| publication_date | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | |
| title | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | /name |
| updated_date | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 | |
| version_label | source · redu unicamp br | connector:redu_unicamp_br@1.0.0 |