Difference between revisions of "Resources for Turkish"
Jump to navigation
Jump to search
Umitmersinli (talk | contribs) |
Umitmersinli (talk | contribs) |
||
Line 29: | Line 29: | ||
* [http://www.hlst.sabanciuniv.edu Sabancı University Natural Language Processing Tools (Turkish Morphological Analyzer, BalkaNET)] | * [http://www.hlst.sabanciuniv.edu Sabancı University Natural Language Processing Tools (Turkish Morphological Analyzer, BalkaNET)] | ||
* [http://ddi.ce.itu.edu.tr Istanbul Technical University Natural Language Processing Research Group] | * [http://ddi.ce.itu.edu.tr Istanbul Technical University Natural Language Processing Research Group] | ||
− | * [http://nooj4nlp.net/turkish Mersin University Turkish National Corpus Project] | + | * [http://nooj4nlp.net/pages/turkish.html NooJ_TR by Mersin University Turkish National Corpus Project Team] |
[[Category:Resources by language|Tajik]] | [[Category:Resources by language|Tajik]] |
Revision as of 14:01, 23 June 2011
Morphological analysis
Free software
- TRMorph "is a relatively complete morphological analyzer for Turkish. It is implemented using SFST, and uses a lexicon based on (but heavily modified) the wordlist of Zemberek spell checker. The morphological analyzer is distributed under the GPL."
Proprietary
Lexical resources
Corpora
Free
- Southeast European Times (sentence aligned corpus, Albanian, Bulgarian, English, Greek, Macedonian, Romanian, Serbo-Croatian, Turkish — approximately 4.5 million words per language)
Proprietary
Bibliography
- K. Oflazer, "Two-level Description of Turkish Morphology," Literary and Linguistic Computing, vol. 9, pp. 137-148, 1995. Backwards PDF