Sampson, Geoffrey (1995) English for the Computer: SUSANNE Corpus and Analytic Scheme. Oxford University Press, USA. ISBN 0198240236Full text not available from this repository.
Computer processing of natural language is a burgeoning field, but until now there has been no agreement on a standardized classification of the diverse structural elements that occur in real-life language material. This book attempts to define a "Linnaean taxonomy" for the English language: an annotation scheme, the SUSANNE scheme, which yields a labelled constituency structure for any string of English, comprehensively identifying all of its surface and logical structural properties. The structure is specified with sufficient rigour that analysts working independently must produce identical annotations for a given example. The scheme is based on large sample of real-life use of British and American written and spoken English. The book also describes the SUSANNE electronic corpus of English which is annotated in accordance with the scheme. It is freely available as a research resource to anyone working at a computer conected to Internet, and since 1992 has come into widespread use in academic and commerical research environments on four continents.
|Schools and Departments:||School of Engineering and Informatics > Informatics|
|Subjects:||Q Science > QA Mathematics > QA0075 Electronic computers. Computer science|
|Depositing User:||Chris Keene|
|Date Deposited:||29 Feb 2008|
|Last Modified:||30 Nov 2012 16:52|
|Google Scholar:||347 Citations|