Knowledge collections and datasets (English)

From ACL Wiki

Revision as of 08:20, 12 December 2006 by Szpektor (talk | contribs)

(diff) ← Older revision | Latest revision (diff) | Newer revision → (diff)

Jump to navigation Jump to search

Datasets for Computational Linguistics and Natural Language Processing.

Clustering by Committee - terms clustered and organized using the Distributional Hypothesis
DIRT Paraphrase Collection - Discovery of Inference Rules from Text
Edinburgh Associative Thesaurus (EAT)
FrameNet
MRC Psycholinguistic Database
Noun Compound Repository
Reuters-21578 Text Categorization Collection
Spam filtering datasets
University of South Florida Free Association Norms
VerbOcean - verbs organized by semantic relation, including temporal precedence and strength
WordNet
WordSimilarity-353 Test Collection
TEASE - Acquisition of Entailment Relations from the Web

Additional Dataset Collections

Linguistic Data Consortium (LDC)

Retrieved from "https://aclweb.org/aclwiki/index.php?title=Knowledge_collections_and_datasets_(English)&oldid=3127"

Knowledge Collections and Datasets