Abstractive Spoken Document Summarization Using Hierarchical Model with Multi-Stage Attention Diversity Optimization

Autor:	Potsawee Manakul, Linlin Wang, Mark J. F. Gales
Rok vydání:	2020
Předmět:	Document summarization business.industry Computer science computer.software_genre multitask learning Hierarchical database model abstractive spoken document summarization Multi stage attention diversity hierarchical model Artificial intelligence business computer Natural language processing Diversity (business)
Zdroj:	INTERSPEECH
DOI:	10.21437/interspeech.2020-1683
Popis:	ive summarization is a standard task for written documents, such as news articles. Applying summarization schemes to spoken documents is more challenging, especially in situations involving human interactions, such as meetings. Here, utterances tend not to form complete sentences and sometimes contain little information. Moreover, speech disfluencies will be present as well as recognition errors for automated systems. For current attention-based sequence-to-sequence summarization systems, these additional challenges can yield a poor attention distribution over the spoken document words and utterances, impacting performance. In this work, we propose a multi-stage method based on a hierarchical encoder-decoder model to explicitly model utterance-level attention distribution at training time; and enforce diversity at inference time using a unigram diversity term. Furthermore, multitask learning tasks including dialogue act classification and extractive summarization are incorporated. The performance of the system is evaluated on the AMI meeting corpus. The inclusion of both training and inference diversity terms improves performance, outperforming current state-of-the-art systems in terms of ROUGE scores. Additionally, the impact of ASR errors, as well as performance on the multitask learning tasks, is evaluated.
Databáze:	OpenAIRE
Externí odkaz:	https://explore.openaire.eu/search/publication?articleId=doi_dedup___::c89a96ecbbc488bc141285613cc6966e https://doi.org/10.21437/interspeech.2020-1683 Zobrazit plný text záznamu