The Linguistic Annotation of Corpora
The TOSCA Analysis System
The article discusses the role of linguistic annotation in corpus linguistics as opposed to annotation in natural language processing. In corpus linguistics, annotation is an integral part of the process of linguistic interpretation and description of the data. Tagging and parsing are discussed as the automatic counterparts of, respectively, the paradigmatic and the syntagmatic description of corpus data. The requirements for a corpus linguistic annotation system are considered. An account is given of the TOSCA analysis system as representative of such an annotation system. Performance results of the system are given, and an evaluation is made.
Keywords: Corpus Linguistics, Annotation, Tagging, Methodology, Parsing
Published online: 01 January 1998
Cited by other publications
Brants, Thorsten, Wojciech Skut & Hans Uszkoreit
This list is based on CrossRef data as of 09 november 2020. Please note that it may not be complete. Sources presented here have been supplied by the respective publishers. Any errors therein should be reported to them.