NST N-gram - Danish News Text

Open data API in a single place

Provided by difi

Get early access to NST N-gram - Danish News Text API!

Let us know and we will figure it out for you.

Dataset information

Catalog

data.norge.no

Country of origin

Norway

Updated

2012.06.11 00:00

Created

2012.01.02

Available languages

Norwegian

Keywords

språkforskning, språkteknologi, ngram, korpus, språkbanken

Datasource

Official portal for European data link

Quality scoring

245

Dataset description

This corpus contains n-grams derived from a 290 million word corpus of Danish news text from the papers Berlingske Tidende, Ekstrabladet og Politiken. The time period covered is 1995-1999. The corpus was originally developed by Nordic Language Technology (NST) 1997-2003. The n-grams were generated by Uni Research for the National Library of Norway. Sequences of one to six words have been generated (i.e., unigrams, bigrams, trigrams, 4-grams, 5-grams and 6-grams) and ordered both by frequency and alphabetically. For convenience, a collection of the 1000 most frequent n-grams of all types listed above is also made available as a separate download.

Build on reliable and scalable technology

FAQ

Frequently Asked Questions

Some basic informations about API Store ®.

Operation and development of APIs are currently fully funded by company Apitalks and its usage is for free.

Yes, you can.

All important information such as time of last update, license and other information are in response of each API call.

In case of major update that would not be compatible with previous version of API, we keep for 30 days both versions so you will have enough time to transfer to new version. We will inform you about the changes in advance by e-mail.

Didn't find the API you need?

Let us know and we will figure it out for you.

API Store ®

API Store provides access to European Open Data via scalable and reliable REST API interface.