Norwegian Parliamentary Speech Corpus 2.0

Open data API in a single place

Provided by Nasjonalbiblioteket

Get early access to Norwegian Parliamentary Speech Corpus 2.0 API!

Let us know and we will figure it out for you.

Dataset information

Country of origin
Updated
2023.07.13 00:00
Created
2019.08.01
Available languages
English
Keywords
tale, korpus, tekst, språkteknologi, språkforskning, språkbanken
Quality scoring

Dataset description

This is version 2.0 of The Norwegian Parliamentary Speech Corpus (NPSC). In version 2.0, a number of changes have been made to the transcriptions, and some identified errors in the corpus have been corrected. The changes are described in detail in the documentation. (Version 1.1 is still available, type "sbr-58" in the search box.) The corpus has been developed by the Norwegian Language Bank at the National Library of Norway from 2019-2021. The NPSC consists of audio recordings of meetings in Stortinget (the Norwegian parliament), with corresponding orthographic transcriptions in either Norwegian Bokmål or Norwegian Nynorsk, as well as various metadata about the speakers. The official proceedings from the meetings are also included in the corpus for reference. The recordings add up to 140 hours of running speech (including pauses) from 267 unique speakers, and contain 65,000 sentences and 1.2 million words in total. Transcription was first done automatically; subsequently, the output of the automatic process was manually checked and corrected by trained linguists and philologists. Finally, all transcriptions were proofread to ensure consistency and accuracy. NPSC is primarily intended as an open-source dataset for ASR development. The individual audio files in the corpus contain the speech of entire days of plenary meetings from 2017 and 2018 (or, if a meeting lasts more than six hours, the first six hours of the meeting). Since the audio files are quite large, individual audio files for each sentence are also included. We greatly appreciate any feedback and suggestions for improvement. Please use our e-mail address, sprakbanken@nb.no.
European data infrastructure with broad catalog discovery, free evaluation access and production-grade API options.
190K+
indexed dataset pages
32
countries and EU institutions
2019
API-first since
Free API quota
for evaluation and prototypes
SLA
history and push on production APIs
FAQ

Questions before production use

Practical answers on evaluation, licensing, freshness, versioning and support.

api.store is built and operated by Apitalks s.r.o. Company details and a direct contact path are linked in the footer for vendor checks and procurement review.
Yes. Selected APIs include a free API quota, so your team can validate coverage, freshness, response shape and workflow fit before asking for a production plan.
Often yes, but usage rights depend on the source license and dataset. We surface source, license and update metadata where available, and can help review terms before a production integration.
Maintained APIs include update metadata where available. For production integrations, we can add history, monitoring and push updates so changes are easier to detect and act on.
Production APIs can add SLA, stable identifiers, versioning support, history, push updates and direct support around the data your product or AI workflow depends on.

Didn't find the API you need?

Let us know and we will figure it out for you.

European data discovery with free evaluation access and production-grade API options.

Copyright © 2026. Made by Apitalks