Grapharizer: A Graph-Based Technique for Extractive Multi-Document Summarization

被引:6
作者
Jalil, Zakia [1 ]
Nasir, Muhammad [2 ]
Alazab, Moutaz [3 ]
Nasir, Jamal [4 ]
Amjad, Tehmina [1 ]
Alqammaz, Abdullah [5 ]
机构
[1] Int Islamic Univ, Dept Comp Sci, Islamabad 44000, Pakistan
[2] Int Islamic Univ, Dept Software Engn, Islamabad 44000, Pakistan
[3] Al Balqa Appl Univ, Fac Artificial Intelligence, Dept Intelligent Syst, Salt 19117, Jordan
[4] Univ Galway, Sch Comp Sci, Galway H91TK33, Ireland
[5] Zarqa Univ, Coll Informat Technol, Dept Cyber Secur, Zarqa 13110, Jordan
关键词
big data; automatic text summarization; extractive multi-document summarization; graph theory; machine learning; anaphora; cataphora; pronoun resolution; grammaticality; topic modeling; ChatGPT; TEXT; SEARCH;
D O I
10.3390/electronics12081895
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Featured Application A graph-based technique tested on a benchmark dataset and augmented by machine learning techniques to provide a concise, informative, and grammatically correct summary. In the age of big data, there is increasing growth of data on the Internet. It becomes frustrating for users to locate the desired data. Therefore, text summarization emerges as a solution to this problem. It summarizes and presents the users with the gist of the provided documents. However, summarizer systems face challenges, such as poor grammaticality, missing important information, and redundancy, particularly in multi-document summarization. This study involves the development of a graph-based extractive generic MDS technique, named Grapharizer (GRAPH-based summARIZER), focusing on resolving these challenges. Grapharizer addresses the grammaticality problems of the summary using lemmatization during pre-processing. Furthermore, synonym mapping, multi-word expression mapping, and anaphora and cataphora resolution, contribute positively to improving the grammaticality of the generated summary. Challenges, such as redundancy and proper coverage of all topics, are dealt with to achieve informativity and representativeness. Grapharizer is a novel approach which can also be used in combination with different machine learning models. The system was tested on DUC 2004 and Recent News Article datasets against various state-of-the-art techniques. Use of Grapharizer with machine learning increased accuracy by up to 23.05% compared with different baseline techniques on ROUGE scores. Expert evaluation of the proposed system indicated the accuracy to be more than 55%.
引用
收藏
页数:26
相关论文
共 50 条
  • [41] A graph-based analytical technique for the improvement of water network model calibration
    Sophocleous, Sophocles
    Savic, Dragan
    Kapelan, Zoran
    Shen, Yibo
    Sage, Paul
    12TH INTERNATIONAL CONFERENCE ON HYDROINFORMATICS (HIC 2016) - SMART WATER FOR THE FUTURE, 2016, 154 : 27 - 35
  • [42] Classification of forensic autopsy reports through conceptual graph-based document representation model
    Mujtaba, Ghulam
    Shuib, Liyana
    Raj, Ram Gopal
    Rajandram, Retnagowri
    Shaikh, Khairunisa
    Al-Garadi, Mohammed Ali
    JOURNAL OF BIOMEDICAL INFORMATICS, 2018, 82 : 88 - 105
  • [43] GCCT: A Graph-Based Coverage and Connectivity Technique for Enhanced Quality of Service in WSN
    Sakkari, Deepak S.
    Basavaraju, T. G.
    WIRELESS PERSONAL COMMUNICATIONS, 2015, 85 (03) : 1295 - 1315
  • [44] GCCT: A Graph-Based Coverage and Connectivity Technique for Enhanced Quality of Service in WSN
    Deepak S. Sakkari
    Basavaraju T. G.
    Wireless Personal Communications, 2015, 85 : 1295 - 1315
  • [45] Graph-Based Multi-Label Classification for WiFi Network Traffic Analysis
    Granato, Giuseppe
    Martino, Alessio
    Baiocchi, Andrea
    Rizzi, Antonello
    APPLIED SCIENCES-BASEL, 2022, 12 (21):
  • [46] A Graph-Based Multi-level Framework to Support the Designing of Collaborative Workplaces
    Di Marino, Castrese
    Rega, Andrea
    Fruggiero, Fabio
    Pasquariello, Agnese
    Vitolo, Ferdinando
    Patalano, Stanislao
    DESIGN TOOLS AND METHODS IN INDUSTRIAL ENGINEERING II, ADM 2021, 2022, : 641 - 649
  • [47] A graph-based sliding window multi-join over data stream
    Byeong-Seob You
    Hae-Young Bae
    重庆邮电大学学报(自然科学版), 2007, (03) : 362 - 366
  • [48] A graph-based sliding window multi-join over data stream
    Zhang Liang
    You, Byeong-Seob
    Ge Jun-wei
    Liu Zhao-hong
    Bae, Hae-Young
    ASGIS 2007: 5TH ASIAN SYMPOSIUM ON GEOGRAPHIC INFORMATION SYSTEMS, 2007, : 362 - 366
  • [49] A review of graph-based multi-agent pathfinding solvers: From classical to classical
    Gao, Jianqi
    Li, Yanjie
    Li, Xinyi
    Yan, Kejian
    Lin, Ke
    Wu, Xinyu
    KNOWLEDGE-BASED SYSTEMS, 2024, 283
  • [50] Deep Program Structure Modeling Through Multi-Relational Graph-based Learning
    Ye, Guixin
    Tang, Zhanyong
    Wang, Huanting
    Fang, Dingyi
    Fang, Jianbin
    Huang, Songfang
    Wang, Zheng
    PACT '20: PROCEEDINGS OF THE ACM INTERNATIONAL CONFERENCE ON PARALLEL ARCHITECTURES AND COMPILATION TECHNIQUES, 2020, : 111 - 123