AMALGUM

AMALGUM is a machine annotated multilayer corpus following the same design and annotation layers as GUM (Georgetown University Multilayer Corpus), but substantially larger (around 4M tokens). The goal of this corpus is to close the gap between high quality, richly annotated, but small datasets, and the larger but shallowly annotated corpora that are often scraped from the Web.

Leave a Reply

Your email address will not be published. Required fields are marked *