Difference between revisions of "Resources for Romanian"

From ACL Wiki
Jump to navigation Jump to search
Line 15: Line 15:
  
 
* [http://www.cs.unt.edu/~rada/downloads.html Romanian NLP]
 
* [http://www.cs.unt.edu/~rada/downloads.html Romanian NLP]
* [http://xixona.dlsi.ua.es/~fran/setimes/ Southeast European Times] (paragraph aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — 9,678 paragraphs, 92,450— 122,912 words per language)
+
* [http://www.statmt.org/setimes/ Southeast European Times] (sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)
 +
 
  
 
===Proprietary===
 
===Proprietary===

Revision as of 12:10, 25 March 2010

Machine translation systems

Free software

Proprietary

Lexical resources

Corpora

Free

  • Romanian NLP
  • Southeast European Times (sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)


Proprietary

  • Corpora (Monolingual, POS tagged and bilingual English/French<->Romanian).

Bibliography

External links