Data · dataset · 2024
Corpus PINO: A spoken language resource for multiple simultaneous comparisons
Listed in IISH Dataverse
Corpus PINO (Corpus Pluristilistico di Italiano e Napoletano Orali, “Multistylistic Corpus of Spoken Italian and Neapolitan”) is a resource designed for research on different styles of spoken Italian and Neapolitan dialect.
Description
The corpus consists of anonymized audio recordings and ELAN time-aligned orthographic transcriptions involving fifty participants (stratified by age, gender, and education level). PINO includes four kinds of spoken activities: sociolinguistic interview; adapted DIAPIX (a “spot the differences” game); reading list; questionnaire with open answer on local language and culture.
Corpus PINO was designed to allow for inter-variety as well as intra-variety analysis. It also allows for analyses of interspeaker variation, or of intra-speaker variation, as each speaker carried out the same four tasks. This structure was thought as a way to encourage systematic and replicable research based on parallel comparisons.
Read the rest (2 more)
The conclusions drawn for the portion of the Italian continuum PINO targets, then, can be used for cross-linguistic comparison with similar continua where quantitative evidence is already available. PINO is also a contribution to the preservation of the local cultural heritage and of a minority language, i.e., an italo-romance dialect. It attests the lives, memories, opinions, traditions, practices, attitudes of fifty members of this community, thus photographing these aspects in a specific moment in time – a post-postmodern society where the tension between global and local plays a pivotal role – and in a place – the province of Naples area – often framed in terms of contradictions, polyvalency, and exceptionality.
Hence, Corpus PINO might be used not only for strictly linguistic or discourse analysis, but for more sociological-based works as well.
Links
Where it is published
- Dataverse dataset page dataverse.nl/dataset.xhtml?persistentId=doi%3A10.34894%2FR1WHEA ↗
landing page · from datasets iisg amsterdam
- DOI doi.org/10.34894/r1whea ↗
DOI / persistent id · from datasets iisg amsterdam
Catalogue records · 1
- Dataverse API dataverse.nl/api/datasets/:persistentId/?persistentId=doi%3A10.34894%2FR1WH… ↗
metadata API · from datasets iisg amsterdam
Topics
- Stated by source
- Arts and Humanities
- From keywords
- Audio · Humanities · Sociolinguistics
Provenance · 1 source records, 10 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| IISH Dataverse | doi:10.34894/R1WHEA | 8 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| concepts[field].anzsrc:field:470411 | mapping · datasets iisg amsterdam | vocabulary-mapper@1.0.0 | keywords['sociolinguistics'] |
| concepts[field].dataverse_subject:arts-and-humanities | source · datasets iisg amsterdam | connector:datasets_iisg_amsterdam@1.0.0 | /subjects |
| concepts[field].local:field:humanities | mapping · datasets iisg amsterdam | connector:datasets_iisg_amsterdam@1.0.0 | /subjects |
| concepts[modality].local:modality:audio | mapping · datasets iisg amsterdam | vocabulary-mapper@1.0.0 | keywords['speech'] |
| created_date | source · datasets iisg amsterdam | connector:datasets_iisg_amsterdam@1.0.0 | |
| description | source · datasets iisg amsterdam | connector:datasets_iisg_amsterdam@1.0.0 | /description |
| publication_date | source · datasets iisg amsterdam | connector:datasets_iisg_amsterdam@1.0.0 | |
| title | source · datasets iisg amsterdam | connector:datasets_iisg_amsterdam@1.0.0 | /name |
| updated_date | source · datasets iisg amsterdam | connector:datasets_iisg_amsterdam@1.0.0 | |
| version_label | source · datasets iisg amsterdam | connector:datasets_iisg_amsterdam@1.0.0 |