e-space
Manchester Metropolitan University's Research Repository

HTSS: A novel hybrid text summarisation and simplification architecture

Zaman, Farooq and Shardlow, Matthew and Hassan, Saeed-Ul and Aljohani, Naif Radi and Nawaz, Raheel (2020) HTSS: A novel hybrid text summarisation and simplification architecture. Information Processing & Management, 57 (6). p. 102351. ISSN 0306-4573

[img]
Restricted to Repository staff only until 13 July 2022.

Download (2MB)

Abstract

Text simplification and text summarisation are related, but different sub-tasks in Natural Language Generation. Whereas summarisation attempts to reduce the length of a document, whilst keeping the original meaning, simplification attempts to reduce the complexity of a document. In this work, we combine both tasks of summarisation and simplification using a novel hybrid architecture of abstractive and extractive summarisation called HTSS. We extend the well-known pointer generator model for the combined task of summarisation and simplification. We have collected our parallel corpus from the simplified summaries written by domain experts published on the science news website EurekaAlert (www.eurekalert.org). Our results show that our proposed HTSS model outperforms neural text simplification (NTS) on SARI score and abstractive text summarisation (ATS) on the ROUGE score. We further introduce a new metric (CSS1) which combines SARI and Rouge and demonstrates that our proposed HTSS model outperforms NTS and ATS on the joint task of simplification and summarisation by 38.94% and 53.40%, respectively. We provide all code, models and corpora to the scientific community for future research at the following URL: https://github.com/slab-itu/HTSS/.

Impact and Reach

Statistics

Downloads
Activity Overview
0Downloads
25Hits

Additional statistics for this dataset are available via IRStats2.

Altmetric

Actions (login required)

Edit Item Edit Item