Dataset Restricted Access
Trial dataset for the SemEval 2019 Task 4: Hyperpartisan News Detection.
The dataset contains 200.000 articles: 100.000 hyperpartisan and 100.000 least biased. All articles are labeled by the overall bias of the publisher as provided by BuzzFeed journalists or MediaBiasFactCheck.com.
The trial data is not fully cleaned. Due to some encoding error, some characters are replaced by question marks. Quote tags are mostly missing. Some text is duplicated. Also, some articles may be contained several times when they are published by several publishers. These errors will be fixed for the final data.
You may request access to the files in this upload, provided that you fulfil the conditions below. The decision whether to grant/deny access is solely under the responsibility of the record owner.
Access is restricted to participants and organizers of the challenge for now. The data will be publicly available after the evaluation period.