Skip to main content

Complex named entities in Spanish texts: Structures and properties

Notice

The full text article is available externally.

View from original source.

We present a linguistic analysis of Named Entities in Spanish texts. Our work is focused on the determination of the structure of complex proper names: names with coordinated constituents, names with prepositional phrases and names formed by several content words initialized by a capital letter. We present the analysis of circa 49,000 examples obtained from Mexican newspapers. We detailed their structure and give some notions about the context surrounding them. Since named entities belong to open class of words they are being created daily, so the challenge for a named entity recognizer is to precisely determine the boundaries of new entity names in any text and to analyze thoroughly their components for deep semantic analysis. Knowing their general classes of structure it should be possible to derive useful heuristics or a specific grammar for natural language processing applications.

Keywords: CONJUNCTIONS; CORPUS LINGUISTICS; DISCOURSE STRUCTURE; NAMED IDENTITY RECOGNITION; NATURAL LANGUAGE PROCESSING; PREPOSITIONS

Document Type: Research Article

Publication date: 01 January 2007

  • Access Key
  • Free content
  • Partial Free content
  • New content
  • Open access content
  • Partial Open access content
  • Subscribed content
  • Partial Subscribed content
  • Free trial content