TRACES Bulgarian Twitter Dataset on Famous Bulgarian Political Cases of Suspected Lies, Annotated with Linguistic Markers of Lies

Open data API in a single place

Provided by Zenodo

Get early access to TRACES Bulgarian Twitter Dataset on Famous Bulgarian Political Cases of Suspected Lies, Annotated with Linguistic Markers of Lies API!

Let us know and we will figure it out for you.

Dataset information

Country of origin
Updated
2024.12.03 00:00
Created
2023.02.07
Available languages
English
Keywords
suspected lies, Twitter, Bulgarian, social media
Quality scoring

Dataset description

This dataset has been created within Project TRACES (more information: https://traces.gate-ai.eu/). The dataset contains 15850 tweet IDs of tweets, written in Bulgarian, with annotations. The dataset can be used for general use or for building lies and disinformation detection applications. Note: this dataset is not fact-checked, the social media messages have been retrieved via keywords. For fact-checked datasets, see our other datasets. The tweets (written between 1 Jan 2020 and 7 July 2022) have been collected via Twitter API under academic access in June-July 2022 with the following keywords without retweets: (ваксиниран депутат) OR (ваксинирани депутати)  (язовири премиер) OR (язовири прокуратура) OR (язовири прокуратурата) ((мвр хемус) OR мвр) (прокуратура OR прокуратурата) (шефът тотото) OR (изпълнителният директор Българския спортен тотализатор) (кирил петков двойно гражданство) OR (премиер двойно гражданство) OR (премиер гражданство) ((Пътна OR загубена OR загуби OR изчезнала) карта газпром) (министър плагиат плагиатство) OR (плагиат плагиатство) ((изслушване главния прокурор) OR (иван гешев))  (фалшива диплома) (златни паспорти) (апартаментгейт OR (къща за гости) OR (къщи за гости) (оръжия OR оръжие) (Украйна OR украина) ((цена OR цени) (газ OR ток OR нафта OR бензин)) (мвр OR данс) (фалшиви новини) (данъци OR данъчни OR данък) ((кораб Царевна) OR Царевна) (Северна Македония)  Explanations of which fields can be used as markers of lies (or of intentional disinformation) are provided in our  paper: Irina Temnikova, Silvia Gargova, Ruslana Margova, Veneta Kireva, Ivo Dzhumerov, Tsvetelina Stefanova and Hristiana Nikolaeva (2023) New Bulgarian Resources for Detecting Disinformation. 10th Language and Technology Conference: Human Language Technologies as a Challenge for Computer Science and Linguistics (LTC'23). Poznań. Poland.
European data infrastructure with broad catalog discovery, free evaluation access and production-grade API options.
190K+
indexed dataset pages
32
countries and EU institutions
2019
API-first since
Free API quota
for evaluation and prototypes
SLA
history and push on production APIs
FAQ

Questions before production use

Practical answers on evaluation, licensing, freshness, versioning and support.

api.store is built and operated by Apitalks s.r.o. Company details and a direct contact path are linked in the footer for vendor checks and procurement review.
Yes. Selected APIs include a free API quota, so your team can validate coverage, freshness, response shape and workflow fit before asking for a production plan.
Often yes, but usage rights depend on the source license and dataset. We surface source, license and update metadata where available, and can help review terms before a production integration.
Maintained APIs include update metadata where available. For production integrations, we can add history, monitoring and push updates so changes are easier to detect and act on.
Production APIs can add SLA, stable identifiers, versioning support, history, push updates and direct support around the data your product or AI workflow depends on.

Didn't find the API you need?

Let us know and we will figure it out for you.

European data discovery with free evaluation access and production-grade API options.

Copyright © 2026. Made by Apitalks