Evaluating Natural Language Processing Systems

Abstract

This report presents a detailed analysis and review of NLP evaluation, in principle and in Practice. Part 1 examines evaluation concepts and establishes a framework for NLP system evaluation. This makes use of experience in the related area of information retrieval and the analysis also refers to evaluation in speech processing. Part 2 surveys significant evaluation work done so far, for instance in machine translation, and discusses the particular problems of generic system evaluation. The conclusion is that evaluation strategies and techniques for NLP need much more development, in particular to take proper account of the influence of system tasks and settings. Part 3 develops a general approach to NLP evaluation, aimed at methodologically-sound strategies for test and evaluation motivated by comprehensive performance factor identification. The analysis throughout the report is supported by extensive illustrative examples.

BibTeX key: Galliers:1993
entry type: techreport
year: 1993
institution: Computer Laboratory, University of Cambridge
number: TR-291
Document: http://citeseer.nj.nec.com/galliers93evaluating.html

BibSonomy

Evaluating Natural Language Processing Systems

Abstract

Tags

Users

Comments and Reviewsshow / hide

Cite this publication

More citation styles

search on