Corpus of Clinical Trials for Evidence-Based-Medicine in Spanish version 3 (CT-EBM-SP v3)

Open data API in a single place

Provided by Agencia Estatal Consejo Superior de Investigaciones Científicas

Get early access to Corpus of Clinical Trials for Evidence-Based-Medicine in Spanish version 3 (CT-EBM-SP v3) API!

Let us know and we will figure it out for you.

Dataset information

Country of origin
Updated
2026.02.04 07:44
Created
2026.02.02
Available languages
English
Keywords
Inter-Annotator Agreement, Inter-Annotator Agreement, Semantic Annotation, Clinical trials, Clinical trials, Semantic Annotation, Natural Language Processing, Natural Language Processing, Evidence-Based Medicine, Evidence-Based Medicine
Quality scoring

Dataset description

This is the version 3 of the CT-EBM-SP corpus of 1200 clinical trials (292173 tokens), annotated with 23 entity types and 18 relation types, covering Unified Medical Language System (UMLS) semantic groups, drug-related information, temporal data, and negation/speculation. It includes 11 encoded attributes (e.g., event temporality and experiencer status) and normalized entities to UMLS Concept Unique Identifiers. The corpus contains 87037 entities, including nested and discontinuous entities, 16597 attributes and 68206 relationships. Inter-annotator agreement (IAA) achieved average F1 values of 0.861 (entities), 0.810 (attributes), and 0.791 (relations). 81.75% of entities were normalized (IAA: F1 = 0.966). The repository includes the code to benchmark this dataset by fine-tuning Transformer models for relation extraction and medical concept normalization. In the relation extraction task, the average F1 ranged from 0.858 to 0.879. In the medical concept normalization task, the accuracy at rank 1 was 0.896.
European data infrastructure with broad catalog discovery, free evaluation access and production-grade API options.
190K+
indexed dataset pages
32
countries and EU institutions
2019
API-first since
Free API quota
for evaluation and prototypes
SLA
history and push on production APIs
FAQ

Questions before production use

Practical answers on evaluation, licensing, freshness, versioning and support.

api.store is built and operated by Apitalks s.r.o. Company details and a direct contact path are linked in the footer for vendor checks and procurement review.
Yes. Selected APIs include a free API quota, so your team can validate coverage, freshness, response shape and workflow fit before asking for a production plan.
Often yes, but usage rights depend on the source license and dataset. We surface source, license and update metadata where available, and can help review terms before a production integration.
Maintained APIs include update metadata where available. For production integrations, we can add history, monitoring and push updates so changes are easier to detect and act on.
Production APIs can add SLA, stable identifiers, versioning support, history, push updates and direct support around the data your product or AI workflow depends on.

Didn't find the API you need?

Let us know and we will figure it out for you.

European data discovery with free evaluation access and production-grade API options.

Copyright © 2026. Made by Apitalks