Translations:Q&A/94/en: Difference between revisions
Appearance
Importing a new version from external source |
(No difference)
|
Latest revision as of 14:19, 5 July 2024
The download files of the Corpus Spoken Dutch (CGN) do not contain the text only. The ort files contain ortographic transcriptions and timestamps and the plk files contain part-of-speech and lemma information. The following perl script takes a list of plk files as input and prints the text. If you run this script from the command line in your terminal, then you can create text files.