Summarization

Thomas Kleinbauer; Gabriel Murray

doi:10.1017/CBO9781139136310.010

10 - Summarization

Published online by Cambridge University Press: 05 July 2012

Thomas Kleinbauer and

Edited by

Jean Carletta and

Thomas Kleinbauer: Affiliation:
Monash University, Australia
Gabriel Murray: Affiliation:
University of British Columbia, Canada
Steve Renals: Affiliation:
University of Edinburgh
Hervé Bourlard: Affiliation:
Idiap Research Institute
Jean Carletta: Affiliation:
University of Edinburgh
Andrei Popescu-Belis: Affiliation:
Idiap Research Institute, Martigny, Switzerland

Book contents

Get access

Summary

Introduction

Automatic summarization has traditionally concerned itself with textual documents. Research on that topic began in the late 1950s with the automatic generation of summaries for technical papers and magazine articles (Luhn, 1958). About 30 years later, summarization has advanced into the field of speech-based summarization, working on dialogues (Kameyama et al., 1996, Reithinger et al., 2000, Alexandersson, 2003) and multi-party interactions (Zechner, 2001b, Murray et al., 2005a, Kleinbauer et al., 2007).

Spärck Jones (1993, 1999) argues that the summarizing process can be described as consisting of three steps (Figure 10.1): interpretation (I), transformation (T), generation (G). The interpretation step analyzes the source, i.e., the input that is to be summarized, and derives from it a representation on which the next step, transformation, operates. In the transformation step the source content is condensed to the most relevant points. The final generation step verbalizes the transformation result into a summary document. This model is a high-level view on the summarization process that abstracts away from the details a concrete implementation has to face.

We distinguish between two general methods for generating an automatic summary: extractive and abstractive. The extractive approach generates a summary by identifying the most salient parts of the source and concatenating these parts to form the actual summary. For generic summarization, these salient parts are sentences that together convey the gist of the document's content. In that sense, extractive summarization becomes a binary decision process for every sentence from the source: should it be part of the summary or not? The selected sentences together then constitute the extractive summary, with some optional post-processing, such as sentence compression, as the final step.

Type: Chapter
Information: Multimodal Signal Processing
Human Interactions in Meetings
, pp. 170 - 192

DOI: https://doi.org/10.1017/CBO9781139136310.010 [Opens in a new window]

Publisher: Cambridge University Press

Print publication year: 2012

Access options

Get access to the full version of this content by using one of the access options below. (Log in options will check for institutional or personal access. Content may require purchase if you do not have access.)

Book contents

10 - Summarization

Summary

Access options

Save book to Kindle

Save book to Dropbox

Save book to Google Drive