Vabamorf - open source morphology tagger for Estonian

Open data API in a single place

Provided by Eesti Keele Instituut

Get early access to Vabamorf - open source morphology tagger for Estonian API!

Let us know and we will figure it out for you.

Dataset information

Country of origin
Updated
2024.05.16 00:00
Created
2023.06.08
Available languages
English
Keywords
[object Object]
Quality scoring

Dataset description

The software package consists of: 1. The morphoanalyzer shall determine the lemma, morphological structure, word type and morphological categories corresponding to the lemma, morphological structure, based on the form of the lemma. If the words are analyzed in many ways, several possible analyses are issued; The analyzer is also able to analyse the words corresponding to which are missing from the basic dictionary of morphological software. By calling out the analytical functions, it is possible to specify whether to analyse the forms of words missing from the basic dictionary (assume the output of the analysis); The analyzer can determine the pronunciation characteristics in the words analyzed: oppressive syllables, letters corresponding to palatalized pronunciations, and third-external syllables. In the case of foreign names, the determination of pronunciation characteristics is not required; 2. The morphosynthesizer synthesizes the morphological structure and pronunciation characteristics of the word change form (s), including the morphological structure and pronunciation characteristics of the change form, on the basis of a predetermined lemma or predetermined change forms. Morphosynthesizer can also synthesize transformation forms from words that are missing from the basic dictionary of morphology software. When summoning synthesis functions, it is possible to specify whether to synthesize the forms of words missing from the basic dictionary (assume the output of synthesis); Synthesis function can synthesize the whole paradigm of other forms at the same time. 3. The spelling module classifies words as orthodox and non-proper word forms. In addition, this module is able to find all the correct language word forms similar to a given word, using a predetermined similarity metric; 4. Estonian language speller suitable for Libre Office and OpenOffice.org 3.0.1 or later. Used on MS Windows, Linux, Macintosh platforms. 5. The Morphological Analysis Connector automatically finds the most likely analysis for multiple analysis words using sentence and/or document context. Preferably, the connector can assign context-dependent probabilities to all the analyses provided for the connection. The connector code shall be separated from the other morphological analysis code. 6. The dictionary. The format of the original text of the dictionary required for the work of the morphological software is human-readable and thoroughly documented. Complementing the basic dictionary of the morphological software by the user can ensure that the dictionary is in conformity with the format (e.g. a dictionary validator can be integrated, which checks whether the dictionary corresponds to the established format and indicates definite and possible errors). The dictionary uses the UTF-8 encoding.
European data infrastructure with broad catalog discovery, free evaluation access and production-grade API options.
190K+
indexed dataset pages
32
countries and EU institutions
2019
API-first since
Free API quota
for evaluation and prototypes
SLA
history and push on production APIs
FAQ

Questions before production use

Practical answers on evaluation, licensing, freshness, versioning and support.

api.store is built and operated by Apitalks s.r.o. Company details and a direct contact path are linked in the footer for vendor checks and procurement review.
Yes. Selected APIs include a free API quota, so your team can validate coverage, freshness, response shape and workflow fit before asking for a production plan.
Often yes, but usage rights depend on the source license and dataset. We surface source, license and update metadata where available, and can help review terms before a production integration.
Maintained APIs include update metadata where available. For production integrations, we can add history, monitoring and push updates so changes are easier to detect and act on.
Production APIs can add SLA, stable identifiers, versioning support, history, push updates and direct support around the data your product or AI workflow depends on.

Didn't find the API you need?

Let us know and we will figure it out for you.

European data discovery with free evaluation access and production-grade API options.

Copyright © 2026. Made by Apitalks