Resources for Bulgarian
From ACLWiki
(Difference between revisions)
(→Machine translation systems) |
|||
| (9 intermediate revisions by 4 users not shown) | |||
| Line 2: | Line 2: | ||
===Free software=== | ===Free software=== | ||
| + | |||
| + | * [https://apertium.svn.sourceforge.net/svnroot/apertium/trunk/apertium-mk-bg apertium-mk-bg] RBMT system between Macedonian and Bulgarian | ||
===Proprietary=== | ===Proprietary=== | ||
| + | |||
| + | * [http://webtrance.skycode.com/?setlang=en WebTrance] | ||
==Lexical resources== | ==Lexical resources== | ||
| + | ===Morphological analysis=== | ||
| + | |||
| + | ====Free software==== | ||
| + | |||
| + | * [https://apertium.svn.sourceforge.net/svnroot/apertium/trunk/apertium-mk-bg/apertium-mk-bg.bg.dix Morphological analyser] 8,581 lemmata, ~88% coverage over SETimes | ||
| + | |||
| + | ====Proprietary==== | ||
| + | |||
| + | == Grammars == | ||
| + | |||
| + | ===Proprietary=== | ||
* [http://dcl.bas.bg/BulNet/general_en.html BulNet WordNet] (21,444 synonym sets) | * [http://dcl.bas.bg/BulNet/general_en.html BulNet WordNet] (21,444 synonym sets) | ||
| + | * [[Generation grammars|KPML generation grammar]] | ||
| + | |||
| + | ==Corpora== | ||
| + | |||
| + | ===Free=== | ||
| + | |||
| + | * [http://www.statmt.org/setimes/ Southeast European Times] (sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language) | ||
| + | |||
| + | ===Proprietary=== | ||
| + | |||
| + | * [http://www.hf.uio.no/easteur-orient/bulg/mat/ Corpus of spoken Bulgarian] | ||
==Bibliography== | ==Bibliography== | ||
| − | |||
==External links== | ==External links== | ||
| − | |||
[[Category:Resources by language|Bulgarian]] | [[Category:Resources by language|Bulgarian]] | ||
Latest revision as of 19:04, 7 October 2010
Contents |
Machine translation systems
Free software
- apertium-mk-bg RBMT system between Macedonian and Bulgarian
Proprietary
Lexical resources
Morphological analysis
Free software
- Morphological analyser 8,581 lemmata, ~88% coverage over SETimes
Proprietary
Grammars
Proprietary
- BulNet WordNet (21,444 synonym sets)
- KPML generation grammar
Corpora
Free
- Southeast European Times (sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)