Constarium
← Search

Data · dataset · 2024

Tweet Annotation Sensitivity Experiment 2

Listed in BERD@NFDI Data Portal

The dataset contains tweet data annotations of hate speech (HS) and offensive language (OL) in five experimental conditions.

Description

The tweet data was sampled from the corpus created by Davidson et al. (2017) . We selected 3,000 Tweets for our annotation.

We developed five experimental conditions that varied the annotation task structure, as shown in the following figure. All tweets were annotated in each condition. Condition A presented the tweet and three options on a single screen: hate speech, offensive language, or neither.

Read the rest (4 more)

Annotators could select one or both of hate speech, offensive language, or indicate that neither applied. Conditions B and C split the annotation of a single tweet across two screens. For Condition B , the first screen prompted the annotator to indicate whether the tweet contained hate speech.

On the following screen, they were shown the tweet again and asked whether it contained offensive language. Condition C was similar to Condition B, but flipped the order of hate speech and offensive language for each tweet. In Conditions D and E, the two tasks are treated independently with annotators being asked to first annotate all tweets for one task, followed by annotating all tweets again for the second task.

Annotators assigned Condition D were first asked to annotate hate speech for all their assigned tweets, and then asked to annotate offensive language for the same set of tweets. Condition E worked the same way, but started with the offensive language annotation task followed by the hate speech annotation task. We recruited US-based annotators from the crowdsourcing platform Prolific during November and December 2022.

Each annotator annotated up to 50 tweets. The dataset also contains demographic information about the annotators. Annotators received a fixed hourly wage in excess of the US federal minimum wage after completing the task.

Links

Where it is published

Catalogue records · 1

Topics

From keywords
Economics & Finance
Inferred from text
Audio 65%
Provenance · 1 source records, 7 field assertions
SourceKeyLast seenRaw
BERD@NFDI Data Portalxvqam-1zf204 d agoJSON v1
FieldAssertionExtractorEvidence
access_levelsource · berd platform deconnector:berd_platform_de@1.0.0
concepts[field].local:field:economics-financemapping · berd platform deconnector:berd_platform_de@1.0.0
concepts[modality].local:modality:audioenrichment · berd platform dekeyword-concept-rules@1.0.0title+description (65%)
descriptionsource · berd platform deconnector:berd_platform_de@1.0.0/metadata/description
publication_datesource · berd platform deconnector:berd_platform_de@1.0.0
titlesource · berd platform deconnector:berd_platform_de@1.0.0/metadata/title
updated_datesource · berd platform deconnector:berd_platform_de@1.0.0