Resources for Finnish
Jump to navigation
Jump to search
Corpora
- Araneum Finnicum, Gigaword Finnish web corpus
- Europarl corpus, sentence aligned with English
- Finnish plain text and Co-occurrences at LCC
- CSC Kielipankki Language Bank at the CSC Scientific Computing Centre, including some 200 million word tokens of Finnish texts.
- HamleDT, harmonized dependency treebanks of many languages, common annotation style.
Morphological analysers
Free software
- Omorfi is an Open Morphology for Finnish, in association with the voikko speller project, see also https://kitwiki.csc.fi/twiki/bin/view/KitWiki/OmorfiHFSTVersion for installing with HFST. (LGPL/GPL)