Deprecated Dataset for "Large Scale Crowdsourcing and Characterization of Twitter Abusive Behavior"

Open data API in a single place

Provided by Zenodo

Get early access to Deprecated Dataset for "Large Scale Crowdsourcing and Characterization of Twitter Abusive Behavior" API!

Let us know and we will figure it out for you.

Dataset information

Country of origin
Updated
2020.03.04 00:00
Created
2018.01.01
Available languages
English
Keywords
Quality scoring

Dataset description

This dataset is deprecated. The updated version of this Dataset is here: https://zenodo.org/record/3678559#.Xl9-Ji97FhE Dataset for the publication "Large Scale Crowdsourcing and Characterization of Twitter Abusive Behavior". Antigoni-Maria Founta, Constantinos Djouvas, Despoina Chatzakou, Ilias Leontiadis, Jeremy Blackburn, Gianluca Stringhini, Athena Vakali, Michael Sirivianos and Nicolas Kourtellis. International AAAI Conference on Web and Social Media (ICWSM), 2018. The dataset provided here includes an updated version of the original dataset, with ~100k tweets annotated using the CrowdFlower platform: hatespeech_labels.csv: contains ~100k rows, where every row consists of a unique Tweet ID and its associated majority annotation UPDATE: It has come to our understanding that a number of the tweets are not available anymore for download on Twitter. Therefore, upon request, we can provide one more file with the full ~100k tweet text and their associated majority labels. The tweets are shuffled so that there is no connection between tweet IDs and texts (in order to be aligned with the T&C of Twitter). To obtain the file contact a.m.founta at gmail dot com AND antonis26papa at gmail dot com. Please cite the paper in any published work that uses any of these resources. @inproceedings{founta2018large,     title={Large Scale Crowdsourcing and Characterization of Twitter Abusive Behavior},     author={Founta, Antigoni-Maria and Djouvas, Constantinos and Chatzakou, Despoina and Leontiadis, Ilias and Blackburn, Jeremy and Stringhini, Gianluca and Vakali, Athena and Sirivianos, Michael and Kourtellis, Nicolas},     booktitle={11th International Conference on Web and Social Media, ICWSM 2018},     year={2018},     organization={AAAI Press} } For any further questions contact a.m.founta at gmail dot com.   Publication DOI: https://doi.org/10.5281/zenodo.1443348 Github: https://github.com/ENCASEH2020/hatespeech-twitter
European data infrastructure with broad catalog discovery, free evaluation access and production-grade API options.
190K+
indexed dataset pages
32
countries and EU institutions
2019
API-first since
Free API quota
for evaluation and prototypes
SLA
history and push on production APIs
FAQ

Questions before production use

Practical answers on evaluation, licensing, freshness, versioning and support.

api.store is built and operated by Apitalks s.r.o. Company details and a direct contact path are linked in the footer for vendor checks and procurement review.
Yes. Selected APIs include a free API quota, so your team can validate coverage, freshness, response shape and workflow fit before asking for a production plan.
Often yes, but usage rights depend on the source license and dataset. We surface source, license and update metadata where available, and can help review terms before a production integration.
Maintained APIs include update metadata where available. For production integrations, we can add history, monitoring and push updates so changes are easier to detect and act on.
Production APIs can add SLA, stable identifiers, versioning support, history, push updates and direct support around the data your product or AI workflow depends on.

Didn't find the API you need?

Let us know and we will figure it out for you.

European data discovery with free evaluation access and production-grade API options.

Copyright © 2026. Made by Apitalks