Extracting information from short messages

  • Authors:
  • Richard Cooper;Sajjad Ali;Chenlan Bi

  • Affiliations:
  • Computing Science, University of Glasgow, Glasgow;Computing Science, University of Glasgow, Glasgow;Computing Science, University of Glasgow, Glasgow

  • Venue:
  • NLDB'05 Proceedings of the 10th international conference on Natural Language Processing and Information Systems
  • Year:
  • 2005

Quantified Score

Hi-index 0.00

Visualization

Abstract

Much currently transmitted information takes the form of e-mails or SMS text messages and so extracting information from such short messages is increasingly important. The words in a message can be partitioned into the syntactic structure, terms from the domain of discourse and the data being transmitted. This paper describes a light-weight Information Extraction component which uses pattern matching to separate the three aspects: the structure is supplied as a template; domain terms are the metadata of a data source (or their synonyms), and data is extracted as those words matching placeholders in the templates.