Improving ROUGE-1 by 6%: A novel multilingual transformer for abstractive news summarization

Citation DataConcurrency and Computation: Practice and Experience, ISSN: 1532-0634, Vol: 36, Issue: 20

Publication Year2024

1
Citations
0
Usage
2
Captures
0
Mentions
0
Social Media

Metric Options: Counts1 Year3 Year

Metrics Details

Citations
1
- Citation Indexes
  1
Captures
2
- Readers
  2

Article Description

Natural language processing (NLP) has undergone a significant transformation, evolving from manually crafted rules to powerful deep learning techniques such as transformers. These advancements have revolutionized various domains including summarization, question answering, and more. Statistical models like hidden Markov models (HMMs) and supervised learning have played crucial roles in laying the foundation for this progress. Recent breakthroughs in transfer learning and the emergence of large-scale models like BERT and GPT have further pushed the boundaries of NLP research. However, news summarization remains a challenging task in NLP, often resulting in factual inaccuracies or the loss of the article's essence. In this study, we propose a novel approach to news summarization utilizing a fine-tuned Transformer architecture pre-trained on Google's mt-small tokenizer. Our model demonstrates significant performance improvements over previous methods on the Inshorts English News dataset, achieving a 6% enhancement in the ROUGE-1 score and reducing training loss by 50%. This breakthrough facilitates the generation of reliable and concise news summaries, thereby enhancing information accessibility and user experience. Additionally, we conduct a comprehensive evaluation of our model's performance using popular metrics such as ROUGE scores, with our proposed model achieving ROUGE-1: 54.6130, ROUGE-2: 31.1543, ROUGE-L: 50.7709, and ROUGE-LSum: 50.7907. Furthermore, we observe a substantial reduction in training and validation losses, underscoring the effectiveness of our proposed approach.

Bibliographic Details

DOI10.1002/cpe.8199

URL IDhttp://www.scopus.com/inward/record.url?partnerID=HzOxMe3b&scp=85195440885&origin=inward; http://dx.doi.org/10.1002/cpe.8199; https://onlinelibrary.wiley.com/doi/10.1002/cpe.8199

AUTHOR(S)

Sandeep Kumar; Arun Solanki

PUBLISHER(S)

Wiley

TAG(S)

Computer Science; Mathematics

Provide Feedback

Have ideas for a new metric? Would you like to see something else here?Let us know