Difference between revisions of "Resources for Bulgarian"
Jump to navigation
Jump to search
(→Free: +Europarl corpus) |
|||
Line 29: | Line 29: | ||
===Free=== | ===Free=== | ||
− | * [http://www.statmt.org/setimes/ Southeast European Times] | + | * [http://www.statmt.org/setimes/ Southeast European Times], sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language |
+ | * [http://www.statmt.org/europarl Europarl corpus], sentence aligned with English | ||
===Proprietary=== | ===Proprietary=== |
Revision as of 10:13, 12 October 2013
Machine translation systems
Free software
- apertium-mk-bg RBMT system between Macedonian and Bulgarian
Proprietary
Lexical resources
Morphological analysis
Free software
- Morphological analyser 8,581 lemmata, ~88% coverage over SETimes
Proprietary
Grammars
Proprietary
- BulNet WordNet (21,444 synonym sets)
- KPML generation grammar
Corpora
Free
- Southeast European Times, sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language
- Europarl corpus, sentence aligned with English