Difference between revisions of "Resources for Finnish"
Jump to navigation
Jump to search
(Added: Araneum) |
|||
(3 intermediate revisions by 3 users not shown) | |||
Line 1: | Line 1: | ||
==Corpora== | ==Corpora== | ||
− | + | * [http://ucts.uniba.sk/aranea_about/ Araneum Finnicum], Gigaword Finnish web corpus | |
+ | * [http://www.statmt.org/europarl Europarl corpus], sentence aligned with English | ||
* [http://corpora.informatik.uni-leipzig.de/ Finnish plain text and Co-occurrences at LCC] | * [http://corpora.informatik.uni-leipzig.de/ Finnish plain text and Co-occurrences at LCC] | ||
* [http://www.csc.fi/english/research/sciences/linguistics/index_html CSC Kielipankki] Language Bank at the [http://www.csc.fi/ CSC] Scientific Computing Centre, including some 200 million word tokens of Finnish texts. | * [http://www.csc.fi/english/research/sciences/linguistics/index_html CSC Kielipankki] Language Bank at the [http://www.csc.fi/ CSC] Scientific Computing Centre, including some 200 million word tokens of Finnish texts. | ||
+ | * [http://ufal.mff.cuni.cz/hamledt HamleDT], harmonized dependency treebanks of many languages, common annotation style. | ||
==Morphological analysers== | ==Morphological analysers== | ||
===Free software=== | ===Free software=== | ||
− | * [https://gna.org/projects/omorfi/ Omorfi], see also https://kitwiki.csc.fi/twiki/bin/view/KitWiki/OmorfiHFSTVersion for installing with [[HFST]]. (LGPL/GPL) | + | * [https://gna.org/projects/omorfi/ Omorfi] is an Open Morphology for Finnish, in association with the [[voikko]] speller project, see also https://kitwiki.csc.fi/twiki/bin/view/KitWiki/OmorfiHFSTVersion for installing with [[HFST]]. (LGPL/GPL) |
[[Category:Resources by language|Finnish]] | [[Category:Resources by language|Finnish]] |
Revision as of 13:10, 8 March 2015
Corpora
- Araneum Finnicum, Gigaword Finnish web corpus
- Europarl corpus, sentence aligned with English
- Finnish plain text and Co-occurrences at LCC
- CSC Kielipankki Language Bank at the CSC Scientific Computing Centre, including some 200 million word tokens of Finnish texts.
- HamleDT, harmonized dependency treebanks of many languages, common annotation style.
Morphological analysers
Free software
- Omorfi is an Open Morphology for Finnish, in association with the voikko speller project, see also https://kitwiki.csc.fi/twiki/bin/view/KitWiki/OmorfiHFSTVersion for installing with HFST. (LGPL/GPL)