Vaccine sentiment analysis using BERT + NBSVM and geo-spatial approaches

© The Author(s) 2023.

Bibliographische Detailangaben
Veröffentlicht in:The Journal of supercomputing. - 1998. - (2023) vom: 07. Mai, Seite 1-31
1. Verfasser: Umair, Areeba (VerfasserIn)
Weitere Verfasser: Masciari, Elio, Ullah, Muhammad Habib
Format: Online-Aufsatz
Sprache:English
Veröffentlicht: 2023
Zugriff auf das übergeordnete Werk:The Journal of supercomputing
Schlagworte:Journal Article Artificial intelligence BERT BERT + NBSVM Buffering COVID vaccines NBSVM Sentiment analysis Spatial analysis Vaccine hesitancy
LEADER 01000caa a22002652c 4500
001 NLM358606845
003 DE-627
005 20250304230122.0
007 cr uuu---uuuuu
008 231226s2023 xx |||||o 00| ||eng c
024 7 |a 10.1007/s11227-023-05319-8  |2 doi 
028 5 2 |a pubmed25n1195.xml 
035 |a (DE-627)NLM358606845 
035 |a (NLM)37359330 
040 |a DE-627  |b ger  |c DE-627  |e rakwb 
041 |a eng 
100 1 |a Umair, Areeba  |e verfasserin  |4 aut 
245 1 0 |a Vaccine sentiment analysis using BERT + NBSVM and geo-spatial approaches 
264 1 |c 2023 
336 |a Text  |b txt  |2 rdacontent 
337 |a ƒaComputermedien  |b c  |2 rdamedia 
338 |a ƒa Online-Ressource  |b cr  |2 rdacarrier 
500 |a Date Revised 28.09.2023 
500 |a published: Print-Electronic 
500 |a Citation Status Publisher 
520 |a © The Author(s) 2023. 
520 |a Since the spread of the coronavirus flu in 2019 (hereafter referred to as COVID-19), millions of people worldwide have been affected by the pandemic, which has significantly impacted our habits in various ways. In order to eradicate the disease, a great help came from unprecedentedly fast vaccines development along with strict preventive measures adoption like lockdown. Thus, world wide provisioning of vaccines was crucial in order to achieve the maximum immunization of population. However, the fast development of vaccines, driven by the urge of limiting the pandemic caused skeptical reactions by a vast amount of population. More specifically, the people's hesitancy in getting vaccinated was an additional obstacle in fighting COVID-19. To ameliorate this scenario, it is important to understand people's sentiments about vaccines in order to take proper actions to better inform the population. As a matter of fact, people continuously update their feelings and sentiments on social media, thus a proper analysis of those opinions is an important challenge for providing proper information to avoid misinformation. More in detail, sentiment analysis (Wankhade et al. in Artif Intell Rev 55(7):5731-5780, 2022. 10.1007/s10462-022-10144-1) is a powerful technique in natural language processing that enables the identification and classification of people feelings (mainly) in text data. It involves the use of machine learning algorithms and other computational techniques to analyze large volumes of text and determine whether they express positive, negative or neutral sentiment. Sentiment analysis is widely used in industries such as marketing, customer service, and healthcare, among others, to gain actionable insights from customer feedback, social media posts, and other forms of unstructured textual data. In this paper, Sentiment Analysis will be used to elaborate on people reaction to COVID-19 vaccines in order to provide useful insights to improve the correct understanding of their correct usage and possible advantages. In this paper, a framework that leverages artificial intelligence (AI) methods is proposed for classifying tweets based on their polarity values. We analyzed Twitter data related to COVID-19 vaccines after the most appropriate pre-processing on them. More specifically, we identified the word-cloud of negative, positive, and neutral words using an artificial intelligence tool to determine the sentiment of tweets. After this pre-processing step, we performed classification using the BERT + NBSVM model to classify people's sentiments about vaccines. The reason for choosing to combine bidirectional encoder representations from transformers (BERT) and Naive Bayes and support vector machine (NBSVM ) can be understood by considering the limitation of BERT-based approaches, which only leverage encoder layers, resulting in lower performance on short texts like the ones used in our analysis. Such a limitation can be ameliorated by using Naive Bayes and Support Vector Machine approaches that are able to achieve higher performance in short text sentiment analysis. Thus, we took advantage of both BERT features and NBSVM features to define a flexible framework for our sentiment analysis goal related to vaccine sentiment identification. Moreover, we enrich our results with spatial analysis of the data by using geo-coding, visualization, and spatial correlation analysis to suggest the most suitable vaccination centers to users based on the sentiment analysis outcomes. In principle, we do not need to implement a distributed architecture to run our experiments as the available public data are not massive. However, we discuss a high-performance architecture that will be used if the collected data scales up dramatically. We compared our approach with the state-of-art methods by comparing most widely used metrics like Accuracy, Precision, Recall and F-measure. The proposed BERT + NBSVM outperformed alternative models by achieving 73% accuracy, 71% precision, 88% recall and 73% F-measure for classification of positive sentiments while 73% accuracy, 71% precision, 74% recall and 73% F-measure for classification of negative sentiments respectively. These promising results will be properly discussed in next sections. The use of artificial intelligence methods and social media analysis can lead to a better understanding of people's reactions and opinions about any trending topic. However, in the case of health-related topics like COVID-19 vaccines, proper sentiment identification could be crucial for implementing public health policies. More in detail, the availability of useful findings on user opinions about vaccines can help policymakers design proper strategies and implement ad-hoc vaccination protocols according to people's feelings, in order to provide better public service. To this end, we leveraged geospatial information to support effective recommendations for vaccination centers 
650 4 |a Journal Article 
650 4 |a Artificial intelligence 
650 4 |a BERT 
650 4 |a BERT + NBSVM 
650 4 |a Buffering 
650 4 |a COVID vaccines 
650 4 |a NBSVM 
650 4 |a Sentiment analysis 
650 4 |a Spatial analysis 
650 4 |a Vaccine hesitancy 
700 1 |a Masciari, Elio  |e verfasserin  |4 aut 
700 1 |a Ullah, Muhammad Habib  |e verfasserin  |4 aut 
773 0 8 |i Enthalten in  |t The Journal of supercomputing  |d 1998  |g (2023) vom: 07. Mai, Seite 1-31  |w (DE-627)NLM098252410  |x 0920-8542  |7 nnas 
773 1 8 |g year:2023  |g day:07  |g month:05  |g pages:1-31 
856 4 0 |u http://dx.doi.org/10.1007/s11227-023-05319-8  |3 Volltext 
912 |a GBV_USEFLAG_A 
912 |a SYSFLAG_A 
912 |a GBV_NLM 
912 |a GBV_ILN_350 
951 |a AR 
952 |j 2023  |b 07  |c 05  |h 1-31