Inicio  /  Algorithms  /  Vol: 16 Par: 11 (2023)  /  Artículo
ARTÍCULO
TITULO

A System to Support Readers in Automatically Acquiring Complete Summarized Information on an Event from Different Sources

Pietro Dell?Oglio    
Alessandro Bondielli and Francesco Marcelloni    

Resumen

Today, most newspapers utilize social media to disseminate news. On the one hand, this results in an overload of related articles for social media users. On the other hand, since social media tends to form echo chambers around their users, different opinions and information may be hidden. Enabling users to access different information (possibly outside of their echo chambers, without the burden of reading entire articles, often containing redundant information) may be a step forward in allowing them to form their own opinions. To address this challenge, we propose a system that integrates Transformer neural models and text summarization models along with decision rules. Given a reference article already read by the user, our system first collects articles related to the same topic from a configurable number of different sources. Then, it identifies and summarizes the information that differs from the reference article and outputs the summary to the user. The core of the system is the sentence classification algorithm, which classifies sentences in the collected articles into three classes based on similarity with the reference article: sentences classified as dissimilar are summarized by using a pre-trained abstractive summarization model. We evaluated the proposed system in two steps. First, we assessed its effectiveness in identifying content differences between the reference article and the related articles by using human judgments obtained through crowdsourcing as ground truth. We obtained an average F1 score of 0.772 against average F1 scores of 0.797 and 0.676 achieved by two state-of-the-art approaches based, respectively, on model tuning and prompt tuning, which require an appropriate tuning phase and, therefore, greater computational effort. Second, we asked a sample of people to evaluate how well the summary generated by the system represents the information that is not present in the article read by the user. The results are extremely encouraging. Finally, we present a use case.

 Artículos similares

       
 
Christogonus U. Onukwube, Daniel O. Aikhuele and Shahryar Sorooshian    
Water distribution networks are complex systems that aid in the delivery of water to residential and non-residential areas. However, the networks can be affected by different types of faults, which could lead to the wastage of treated water. As such, the... ver más
Revista: Applied Sciences

 
Shweta More, Moad Idrissi, Haitham Mahmoud and A. Taufiq Asyhari    
The rapid proliferation of new technologies such as Internet of Things (IoT), cloud computing, virtualization, and smart devices has led to a massive annual production of over 400 zettabytes of network traffic data. As a result, it is crucial for compani... ver más
Revista: Algorithms

 
Beichen Lu, Yanjun Liu, Xiaoyu Zhai, Li Zhang and Yun Chen    
In recent years, clean and renewable energy sources have received much attention to balance the contradiction between resource needs and environmental sustainability. Among them, ocean thermal energy conversion (OTEC), which consists of surface warm seaw... ver más

 
Chi-Hieu Ngo, Seok-Ju Lee, Changhyun Kim, Minh-Chau Dinh and Minwon Park    
In seaports, the automatic Grab-Type Ship Unloader (GTSU) stands out for its ability to automatically load and unload materials, offering the potential for substantial productivity improvement and cost reduction. Developing a fully automatic GTSU, howeve... ver más

 
Mario E. Rivero-Angeles, Iclia Villordo-Jimenez, Izlian Y. Orea-Flores, Noé Torres-Cruz and Angel Pretelín Ricárdez    
In modern and future communication systems, we expect peaks of traffic that largely exceed the capacity of the system, since they are originally designed to support normal traffic loads. Such peaks can be caused by emergency events and cultural or sporti... ver más
Revista: Information