Translations:Parallel Monolingual Corpora/2/en

From Clarin K-Centre
Jump to navigation Jump to search

The DAESO Corpus is a parallel monolingual treebank of Dutch texts and the corpus contains more than 2.1 million words of parallel and comparable text. About 678,000 words were lined up manually and about 1.5 million words were automatically aligned. A semantic relation was added to the aligned words / phrases.