Search CORE

1,492 research outputs found

Syntactic annotation of non-canonical linguistic structures

Author: Doolittle Seanna
Hirschmann Hagen
Lüdeling Anke
Publication venue
Publication date: 27/10/2009
Field of study

This paper deals with the syntactic annotation of corpora that contain both ‘canonical’ and ‘non-canonical’ sentences

Hochschulschriftenserver - Universität Frankfurt am Main

Information structure

Author: Endriss Cornelia
Fiedler Ines
Götze Michael
Hinterwimmer Stefan
Petrova Svetlana
Schwarz Anne
Skopeteas Stavros
Stoel Ruben
Weskott Thomas
Publication venue
Publication date: 05/05/2009
Field of study

The guidelines for Information Structure include instructions for the annotation of Information Status (or ‘givenness’), Topic, and Focus, building upon a basic syntactic annotation of nominal phrases and sentences. A procedure for the annotation of these features is proposed

Hochschulschriftenserver - Universität Frankfurt am Main

Modelling Discourse-related terminology in OntoLingAnnot’s ontologies

Author: Aguado de Cea G.
Pareja-Lora A.
Publication venue: Facultad de Informática (UPM)
Publication date: 01/01/2010
Field of study

Recently, computational linguists have shown great interest in discourse annotation in an attempt to capture the internal relations in texts. With this aim, we have formalized the linguistic knowledge associated to discourse into different linguistic ontologies. In this paper, we present the most prominent discourse-related terms and concepts included in the ontologies of the OntoLingAnnot annotation model. They show the different units, values, attributes, relations, layers and strata included in the discourse annotation level of the OntoLingAnnot model, within which these ontologies are included, used and evaluated

Archivo Digital UPM

RDF/S)XML Linguistic Annotation of Semantic Web Pages

Author: Aguado de Cea G.
Pareja-Lora A.
Plaza Arteche R.
Álvarez de Mon Rego I.
Publication venue: Facultad de Informática (UPM)
Publication date: 01/01/2002
Field of study

Although with the Semantic Web initiative much research on web pages semantic annotation has already done by AI researchers, linguistic text annotation, including the semantic one, was originally developed in Corpus Linguistics and its results have been somehow neglected by AI. ..

Archivo Digital UPM

What linguists always wanted to know about german and did not know how to estimate

Author: Hinrichs Erhard
Kübler Sandra
Publication venue
Publication date: 01/01/2006
Field of study

This paper profiles significant differences in syntactic distribution and differences in word class frequencies for two treebanks of spoken and written German: the TüBa-D/S, a treebank of transliterated spontaneous dialogues, and the TüBa-D/Z treebank of newspaper articles published in the German daily newspaper die tageszeitung´(taz). The approach can be used more generally as a means of distinguishing and classifying language corpora of different genres

Hochschulschriftenserver - Universität Frankfurt am Main