Kernels for text

John Shawe-Taylor; Nello Cristianini

doi:10.1017/CBO9780511809682.011

10 - Kernels for text

from Part III - Constructing kernels

Published online by Cambridge University Press: 29 March 2011

John Shawe-Taylor and

Nello Cristianini

Show author details

John Shawe-Taylor: Affiliation:
University of Southampton
Nello Cristianini: Affiliation:
University of California, Davis

Book contents

Get access

Summary

The last decade has seen an explosion of readily available digital text that has rendered attempts to analyse and classify by hand infeasible. As a result automatic processing of natural language text documents has become a main research interest of Artificial Intelligence (AI) and computer science in general. It is probably fair to say that after multivariate data, natural language text is the most important data format for applications. Its particular characteristics therefore deserve specific attention.

We will see how well-known techniques from Information Retrieval (IR), such as the rich class of vector space models, can be naturally reinterpreted as kernel methods. This new perspective enriches our understanding of the approach, as well as leading naturally to further extensions and improvements. The approach that this perspective suggests is based on detecting and exploiting statistical patterns of words in the documents. An important property of the vector space representation is that the primal–dual dialectic we have developed through this book has an interesting counterpart in the interplay between term-based and document-based representations.

The goal of this chapter is to introduce the Vector Space family of kernel methods highlighting their construction and the primal–dual dichotomy that they illustrate. Other kernel constructions can be applied to text, for example using probabilistic generative models and string matching, but since these kernels are not specific to natural language text, they will be discussed separately in Chapters 11 and 12.

Type: Chapter
Information: Kernel Methods for Pattern Analysis , pp. 327 - 343

DOI: https://doi.org/10.1017/CBO9780511809682.011 [Opens in a new window]

Publisher: Cambridge University Press

Print publication year: 2004

Access options

Get access to the full version of this content by using one of the access options below. (Log in options will check for institutional or personal access. Content may require purchase if you do not have access.)

Book contents

10 - Kernels for text

Summary

Access options

Save book to Kindle

Save book to Dropbox

Save book to Google Drive