Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

EXPLORING ABSTRACTIVE SUMMARIZATION OF PRE-TRAINED MODELS: A STUDY ON GPT-2, T5, PEGASUS AND BART

View through CrossRef
Text summarization, a significant challenge within Natural Language Processing (NLP), purposes to distill huge volumes of content into brief, coherent summaries. As the amount of text data continues to expand, the need for efficient and accurate summarization methods has grown, making it a critical task across various sectors. Despite advancements in summarization, especially with the emergence of transfer learning that leverage the capabilities of pre-trained models, questions remain about which models perform best for summarization on specific datasets. This study evaluates how well the pre-trained models—GPT-2, T5, Pegasus, and BART perform for the abstractive text summarization task. We employ the datasets MultiNews, WikiSum, DUC & CNN/Daily Mail, consisting of news related articles accompanied by their human-generated summaries. In the proposed work, we fine-tuned each model on these datasets through transfer learning, by carefully adjusting the parameters. The metrics employed to assess the performance of each model are ROUGE and METEOR. After rigorous experiments, the results indicate that the T5 model beats the others in abstractive summarization on these datasets, achieving superior ROUGE scores of R1, R2 and RL equal to 63.04%, 42.22%, 47.82% respectively and an average precision of 90.10%. Lastly, the further exploration of future research directions is provided.
Title: EXPLORING ABSTRACTIVE SUMMARIZATION OF PRE-TRAINED MODELS: A STUDY ON GPT-2, T5, PEGASUS AND BART
Description:
Text summarization, a significant challenge within Natural Language Processing (NLP), purposes to distill huge volumes of content into brief, coherent summaries.
As the amount of text data continues to expand, the need for efficient and accurate summarization methods has grown, making it a critical task across various sectors.
Despite advancements in summarization, especially with the emergence of transfer learning that leverage the capabilities of pre-trained models, questions remain about which models perform best for summarization on specific datasets.
This study evaluates how well the pre-trained models—GPT-2, T5, Pegasus, and BART perform for the abstractive text summarization task.
We employ the datasets MultiNews, WikiSum, DUC & CNN/Daily Mail, consisting of news related articles accompanied by their human-generated summaries.
In the proposed work, we fine-tuned each model on these datasets through transfer learning, by carefully adjusting the parameters.
The metrics employed to assess the performance of each model are ROUGE and METEOR.
After rigorous experiments, the results indicate that the T5 model beats the others in abstractive summarization on these datasets, achieving superior ROUGE scores of R1, R2 and RL equal to 63.
04%, 42.
22%, 47.
82% respectively and an average precision of 90.
10%.
Lastly, the further exploration of future research directions is provided.

Related Results

A Comprehensive Study of Text Summarization with Advent of Large Language Models
A Comprehensive Study of Text Summarization with Advent of Large Language Models
Introduction: Communication is at the heart of the human race. With the growth of social media and other communication platforms, the globe is now connected at a single click. Peop...
Automatic summarization of Malayalam documents using clause identification method
Automatic summarization of Malayalam documents using clause identification method
<span>Text summarization is an active research area in the field of natural language processing. Huge amount of information in the internet necessitates the development of au...
Performance of Novel GPT-4 in Otolaryngology Knowledge Assessment
Performance of Novel GPT-4 in Otolaryngology Knowledge Assessment
Abstract Purpose GPT-4, recently released by OpenAI, improves upon GPT-3.5 with increased reliability and expanded capabilities, including user-spec...
Abstractive and Extractive Approaches for Summarizing Multi-document Travel Reviews
Abstractive and Extractive Approaches for Summarizing Multi-document Travel Reviews
Travel reviews offer insights into users' experiences at places they have visited, including hotels, restaurants, and tourist attractions. Reviews are a type of multidocument, wher...
Text summarization: BART, RF, and hybrid BART-RF algorithm comparison
Text summarization: BART, RF, and hybrid BART-RF algorithm comparison
Data and information accumulate quantitatively and qualitatively. Abundant text data are posted on the internet. The number correlates to the complexity of the summarization. Autom...
Placenta-derived Extracellular Vesicles in Maternal Plasma of Hb Bart’s Fetuses
Placenta-derived Extracellular Vesicles in Maternal Plasma of Hb Bart’s Fetuses
Abstract Introduction: Alpha-thalassemia is the most common cause of hydrops fetalis among Southeast Asians (also called “Bart’...

Back to Top