Social Media Ban for Minors: A Computational Analysis of Media Coverage in Europe and Beyond, Dataset

Open data API in a single place

Provided by European Commission, Joint Research Centre

Get early access to Social Media Ban for Minors: A Computational Analysis of Media Coverage in Europe and Beyond, Dataset API!

Let us know and we will figure it out for you.

Dataset information

Country of origin
Updated
2025.12.08 00:00
Created
2025.12.08
Available languages
English
Keywords
multilingual clustering, media coverage, persuasion techniques, framing dimensions, social media
Quality scoring

Dataset description

This dataset comprises more than 10,000 news articles referring to social media bans for minors, published between 1 January 2024 and 15 March 2025. It served as the primary input for the publication “Social Media Ban for Minors: A Computational Analysis of Media Coverage in Europe and Beyond.” The data are organised by month. Due to data sensitivity considerations, we are unable to provide a dataset containing a separate list of unverified online sources. For this reason, all sources, both mainstream and unverified, are included together within one dataset. The identification of unverified sources is based on assessments conducted by the European External Action Service (EEAS), as outlined in the “3rd EEAS Report on Foreign Information Manipulation and Interference Threats”, as well as evaluations made by independent external experts working in the field of disinformation. Some of these sources are reviewed by reputable fact-checking organisations such as mediascan, butac, factcheck, cbsnews, and konspiratori. For each article, the following fields are provided: Title, Link, Publication Date, Source, Guid, Cluster Title, Cluster Keyphrases, Cluster Summary, Framing Dimensions, and Persuasion Techniques. By tracking coverage over time and applying multilingual clustering combined with large language model (LLM)–based cluster summarisation, analysts identified the key narratives surrounding debates on social media bans. The “cluster” field is derived from a multilingual clustering pipeline using the LaBSE sentence embedding model, PyNNDescent for approximate neighbourhood graphs, and LeidenAlg for community detection. Each cluster represents a story or narrative prominent in a given month. For each cluster, 100 random article excerpts (first 350 characters) were sampled, and GPT-4, GPT-4-turbo, and GPT-3.5-turbo were used to generate a cluster title, cluster keyphrases, and a cluster summary. The “framing dimensions” and “persuasion techniques” fields contain specific frames and rhetorical strategies identified within each article. Articles may contain multiple instances. These labels were produced using in-house machine-learning classifiers. Framing refers to the perspective under which an issue or a piece of news is presented. We consider 14 frames: (1) Economic, (2) Capacity and resources, (3) Morality, (4) Fairness and equality, (5) Legality, constitutionality and jurisprudence, (6) Policy prescription and evaluation, (7) Crime and punishment, (8) Security and defence, (9) Health and safety, (10) Quality of life, (11) Cultural identity, (12) Public opinion, (13) Political, (14) External regulation and reputation . Persuasion techniques refer to the style of writing of a text with the aim to influence the reader. In this report we consider the following sub selection: (1) Appeal to Authority, (2) Appeal to Fear-Prejudice, (3) Appeal to Hypocrisy, (4) Appeal to Time, (5) Appeal to Values, (6) Causal Oversimplification, (7) Consequential Oversimplification, (8) Conversation Killer, (9) Doubt, (10) Exaggeration-Minimisation, (11) False Dilemma-No Choice, (12) Flag Waving, (13) Guilt by Association, (14) Loaded Language, (15) Name Calling-Labelling, (16) Questioning the Reputation, (17) Repetition, (18) Slogan.
European data infrastructure with broad catalog discovery, free evaluation access and production-grade API options.
190K+
indexed dataset pages
32
countries and EU institutions
2019
API-first since
Free API quota
for evaluation and prototypes
SLA
history and push on production APIs
FAQ

Questions before production use

Practical answers on evaluation, licensing, freshness, versioning and support.

api.store is built and operated by Apitalks s.r.o. Company details and a direct contact path are linked in the footer for vendor checks and procurement review.
Yes. Selected APIs include a free API quota, so your team can validate coverage, freshness, response shape and workflow fit before asking for a production plan.
Often yes, but usage rights depend on the source license and dataset. We surface source, license and update metadata where available, and can help review terms before a production integration.
Maintained APIs include update metadata where available. For production integrations, we can add history, monitoring and push updates so changes are easier to detect and act on.
Production APIs can add SLA, stable identifiers, versioning support, history, push updates and direct support around the data your product or AI workflow depends on.

Didn't find the API you need?

Let us know and we will figure it out for you.

European data discovery with free evaluation access and production-grade API options.

Copyright © 2026. Made by Apitalks