Difference between revisions of "Resources for Albanian"

From ACL Wiki
Jump to: navigation, search
(setimes)
 
 
(2 intermediate revisions by 2 users not shown)
Line 5: Line 5:
 
===Proprietary===
 
===Proprietary===
  
 +
 +
==Morphological analysis==
 +
 +
===Free software===
 +
 +
* [https://apertium.svn.sourceforge.net/svnroot/apertium/incubator/apertium-mk-sq/sq.xml Morphological analyser] 3,308 lemmata, ~80% coverage over SETimes
 +
 +
===Proprietary===
  
 
==Corpora==
 
==Corpora==
Line 10: Line 18:
 
===Free===
 
===Free===
  
* [http://xixona.dlsi.ua.es/~fran/setimes/ Southeast European Times] (paragraph aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — 9,678 paragraphs, 92,450— 122,912 words per language)
+
* [http://www.statmt.org/setimes/ Southeast European Times] (paragraph aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)
  
 
==Bibliography==
 
==Bibliography==

Latest revision as of 15:59, 7 October 2010

Machine translation systems

Free software

Proprietary

Morphological analysis

Free software

Proprietary

Corpora

Free

  • Southeast European Times (paragraph aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)

Bibliography

External links