Constarium
← Search

Data · dataset · 2014

Mining cross-domain rating datasets from structured data on twitter

Listed in BERD@NFDI Data Portal

While rating data is essential for all recommender systems research, there are only a few public rating datasets available, most of them years old and limited to the movie domain.

Description

With this work, we aim to end the lack of rating data by illustrating how vast amounts of ratings can be unambiguously collected from Twitter. We validate our approach by mining ratings from four major online websites focusing on movies, books, music and video clips.

In a short mining period of 2 weeks, close to 3 million ratings were collected. Since some users turned up in more than one dataset, we believe this work to be amongst the first to provide a true cross-domain rating dataset

Links

Where it is published

Catalogue records · 1

Topics

From keywords
Economics & Finance
Inferred from text
Econometrics 69% · Video 75%
Provenance · 1 source records, 8 field assertions
SourceKeyLast seenRaw
BERD@NFDI Data Portalj598z-nhe039 d agoJSON v1
FieldAssertionExtractorEvidence
access_levelsource · berd platform deconnector:berd_platform_de@1.0.0
concepts[field].anzsrc:group:3802enrichment · berd platform detaxonomy-embedding@1.0.0title+keywords+description (69%)
concepts[field].local:field:economics-financemapping · berd platform deconnector:berd_platform_de@1.0.0
concepts[modality].local:modality:videoenrichment · berd platform dekeyword-concept-rules@1.0.0title+description (75%)
descriptionsource · berd platform deconnector:berd_platform_de@1.0.0/metadata/description
publication_datesource · berd platform deconnector:berd_platform_de@1.0.0
titlesource · berd platform deconnector:berd_platform_de@1.0.0/metadata/title
updated_datesource · berd platform deconnector:berd_platform_de@1.0.0