Constarium
← Search

Data · dataset · 2026

Ethnic Stereotypes in Language: A Validated Sentence–Word Priming Stimulus Database for EEG/Neuroimaging, Behavioral, and LLM Research

Listed in Bicocca Open Archive Research Data

This database provides a validated set of sentences designed for research on ethnic stereotyping and social cognition.

Description

The materials are suitable for use in EEG/ERP experiments, neuroimaging studies (fMRI, MEG), behavioral paradigms, and computational/LLM modelling of social bias. The stimuli were validated on a native Italian-speaking population.

The database consists of 330 sentence–word pairs organized across three experimental conditions: • Incongruent (N = 110): sentences in which the terminal target word violates the ethnic stereotype associated with the speaker's accent (e.g., a speaker with an African accent described as a pilot or a surgeon) • Congruent (N = 110): sentences in which the terminal word is consistent with prevailing ethnic stereotypes (e.g., a speaker with an African accent described as working in a field or selling items at a market stall) • Neutral (N = 110): control sentences mentioning Italian individuals or regional groups, with no ethnic accent or stereotype content Sentences are distributed across seven ethnic/accent categories: African, Chinese, Roma/Sinti, Eastern European, Asian, Latin American, and Arabian.

Read the rest (2 more)

All three conditions are matched on the following variables: • Sentence length in characters • Sentence length in words • Terminal word length in characters • Lexical frequency of the terminal word • Grammatical category and imageability of the terminal word • Gender of the protagonist (male / female / mixed) Intended uses • EEG/ERP studies of stereotyping, social prediction, and language processing (N400, LAN, LP components) • Neuroimaging (fMRI/MEG) studies of social cognition and predictive coding • Behavioral studies of implicit ethnic bias, semantic priming, and accent perception • Training, testing, or probing Large Language Models (LLMs) on social bias, world knowledge and stereotype content Availability of auditory stimuli A professionally recorded auditory version of these stimuli (sentences spoken with native-accented voices across the seven ethnic/accent categories) is also available.

Auditory stimuli are not shared openly, but may be made available to research groups for non-profit academic purposes under a co-authorship agreement: groups wishing to use the auditory materials in a study intended for publication are invited to contact the authors to discuss the terms of scientific collaboration, which will include co-authorship on any resulting paper. To inquire, write to: mado.proverbio@unimib.it

Links

Where it is published

Catalogue records · 1

Topics

Provenance · 1 source records, 10 field assertions
SourceKeyLast seenRaw
Bicocca Open Archive Research Dataoai:board.unimib.it/nbwn3jwzh4.15 d agoJSON v1
FieldAssertionExtractorEvidence
concepts[field].anzsrc:group:5204enrichment · board unimib ittaxonomy-embedding@1.0.0title+keywords+description (75%)
concepts[field].local:field:earth-environmentalmapping · board unimib itconnector:board_unimib_it@1.0.0
concepts[field].local:field:engineeringmapping · board unimib itconnector:board_unimib_it@1.0.0
concepts[field].local:field:humanitiesmapping · board unimib itconnector:board_unimib_it@1.0.0
concepts[field].local:field:life-sciencesmapping · board unimib itconnector:board_unimib_it@1.0.0
concepts[field].local:field:psychology-behavioralmapping · board unimib itconnector:board_unimib_it@1.0.0
concepts[field].local:field:social-sciencemapping · board unimib itconnector:board_unimib_it@1.0.0
descriptionsource · board unimib itconnector:board_unimib_it@1.0.0/metadata/dc/description
publication_datesource · board unimib itconnector:board_unimib_it@1.0.0
titlesource · board unimib itconnector:board_unimib_it@1.0.0/metadata/dc/title