000 03889ntm a2200337 i 4500
999 _c100883
_d100889
003 MY-KuUP
005 20251125110901.0
006 a||||fr|||| 00| 0
007 ta
008 240530t20232023my a|||fs|||| 000 0 eng d
020 _aTHE0009866 (Local)
_qHardback
040 _aUMPSA
_beng
_cUMP
_erda
090 _aPSM .Q57 2023 r Bc
100 1 _aQiryn Adriana Binti Kharul Zaman,
_eauthor.
245 1 0 _aPredictive analytics for the sentiment of malaysian place of interest using machine learning models /
_cQiryn Adriana Binti Kharul Zaman
264 1 _aKuantan, Pahang:
_bUMPSA,
_c2023
264 4 _c©2023
300 _axiv, 92 pages :
_bIllustration (some colour) ;
_e1 CD-ROM.
336 _2rdacontent
_atext
337 _2rdamedia
_aunmediated
338 _2rdacarrier
_avolume
347 _2rda
_atext file
_bPDF
500 _aCenter for Mathematical Sciences
502 _aBachelor of Applied Science in Data Analytics with Honours--Universiti Malaysia Pahang – 2023
504 _aIncludes bibliographical reference
520 3 _aSentiment analysis is a method of automatically identifying sentiments expressed in online interactions with the aim of evaluating users or customers' opinions on a product, brand, or service. It helps companies gain valuable insights and respond to their customers more efficiently. As people use social media, forums, blogs, and the web to express their opinions on various discussion topics, these channels have become an ideal domain for utilizing customer sentiment analysis. The focus of this study is to conduct Natural Language Processing (NLP) on tweets and make a better classification of sentiment using Malaya. Furthermore, this study also trains three machine learning algorithms to predict the sentiment of textual data. Moreover, this study also creates dashboard to visualize social media insights and suggest recommendations based on the insights. The study gathered users or customers feedback from Twitter on 1st January 2023 to 1st March 2023 using the social media monitoring software, Determ, which retrieves tweets in real-time based on search terms, time, users, and likes. The tweets containing feedbacks and responses were organized into tables and saved as a CSV file. Subsequently, the study proceeded with the pre-processing stage to handle missing and erroneous values in the data. Additionally, several Natural Language Processing (NLP) techniques were employed to pre-process the text data, as part of the machine learning process. The data was then divided into training and testing sets, and was trained using three different supervised learning algorithms, namely Support Vector Machine, Random Forest, and Naive Bayes. Finally, the performance of prediction of each model was compared to identify the most accurate one, and based on the analysis, it was concluded that Support Vector Machine exhibited the best performance in terms of accuracy, recall score, F1 score, and precision score. Furthermore, it is worth noting that this sentiment analysis research is extended to analyze sentiments expressed in texts written in Malay language by utilizing the Natural Language-Toolkit library for Bahasa Malaysia, powered by Tensorflow and PyTorch. Regarding the outcomes of the customer sentiment analysis, there are some recommendations that can be adopted to enhance the effectiveness of the study. For instance, the analysis can be extended to include customer feedback data collected from social media platforms such as Facebook, Instagram, Tik Tok, web and forums. Additionally, the performance of the customer sentiment analysis model can be improved by leveraging deep learning techniques.
610 2 0 _aCenter for Mathematical Sciences
_xDissertations
650 0 _aUniversities and colleges
_xDissertations
650 0 _aFinal Year Project
942 _2lcc
_cPSM