A position-aware language modeling framework for Extractive broadcast news speech summarization
Journal
ACM Transactions on Asian and Low-Resource Language Information Processing
Journal Volume
16
Journal Issue
4
Date Issued
2017
Author(s)
Abstract
Extractive summarization, a process that automatically picks exemplary sentences from a text (or spoken) document with the goal of concisely conveying key information therein, has seen a surge of attention from scholars and practitioners recently. Using a language modeling (LM) approach for sentence selection has been proven effective for performing unsupervised extractive summarization. However, one of the major difficulties facing the LM approach is to model sentences and estimate their parameters more accurately for each text (or spoken) document. We extend this line of research and make the following contributions in this work. First, we propose a position-aware language modeling framework using various granularities of position-specific information to better estimate the sentence models involved in the summarization process. Second, we explore disparate ways to integrate the positional cues into relevance models through a pseudo-relevance feedback procedure. Third, we extensively evaluate various models originated from our proposed framework and several well-established unsupervised methods. Empirical evaluation conducted on a broadcast news summarization task further demonstrates performance merits of the proposed summarization methods. © 2017 ACM.
Subjects
Extractive summarization; Positional language modeling; Relevance modeling; Speech information
SDGs
Other Subjects
Computational linguistics; Natural language processing systems; Empirical evaluations; Extractive summarizations; Language model; Position-specific information; Pseudo relevance feedback; Relevance models; Speech information; Speech summarization; Modeling languages
Type
journal article
