Automatic creation of bilingual dictionaries for Finno-Ugric languages

We introduce an ongoing project whose objective is to provide linguistically based support for several small Finno-Ugric digital communities in generating online content. To achieve our goals, we collect parallel, comparable and monolingual text material for the following Finno-Ugric (FU) languages:...

Full description

Bibliographic Details
Published in:Septentrio Conference Series
Main Authors: Simon, Eszter, Benyeda, Ivett Zs., Koczka, Péter, Ludányi, Zsófia
Format: Article in Journal/Newspaper
Language:English
Published: Septentrio Academic Publishing 2015
Subjects:
Online Access:https://septentrio.uit.no/index.php/SCS/article/view/3474
https://doi.org/10.7557/5.3474
Description
Summary:We introduce an ongoing project whose objective is to provide linguistically based support for several small Finno-Ugric digital communities in generating online content. To achieve our goals, we collect parallel, comparable and monolingual text material for the following Finno-Ugric (FU) languages: Komi-Zyrian and Permyak, Udmurt, Meadow and Hill Mari and Northern Sami, as well as for major languages that are of interest to the FU community: English, Russian, Finnish and Hungarian. Our goal is to generate proto-dictionaries for the mentioned language pairs and deploy the enriched lexical material on the web in the framework of the collaborative dictionary project Wiktionary. In addition, we will make all of the project’s products (corpora, models, dictionaries) freely available supporting further research.