Claim Missing Document
Check
Articles

Enhancing Review Processing in the Video Game Adaptation Domain through VADER and Rating-Based Labeling using SVM Sajmira, Danita Divka; Umam, Khothibul; Handayani, Maya Rini
Jurnal Sisfokom (Sistem Informasi dan Komputer) Vol. 14 No. 3 (2025): JULY
Publisher : ISB Atma Luhur

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.32736/sisfokom.v14i3.2409

Abstract

The adaptation of video games into films or television series has increasingly become a prominent trend in the entertainment sector, often eliciting diverse reactions from audiences.A prime example is The Last of Us, a video game adaptation series that generated substantial online discussions and sentiment, and serves as the specific case study in this research. Sentiment patterns found in audience reviews of The Last of Us on IMDb are analyzed using a domain-specific classification framework tailored to the language characteristics of entertainment media. A key issue addressed is the discrepancy between numerical ratings and the sentiment conveyed in review texts, which may lead to inconsistent labeling. The study employs a machine learning technique, Support Vector Machine (SVM), coupled with two distinct labeling methods: manual labeling based on IMDb ratings, and automatic labeling using the lexicon-driven VADER tool. A total of 2,017 English reviews of The Last of Us were gathered via web scraping from IMDb, followed by preprocessing, TF-IDF feature extraction, and hyperparameter optimization using RandomizedSearchCV. These results show that the SVM model trained on VADER-labeled data achieved an accuracy of 0.97, outperforming the model trained on manually labeled data at 0.79. Lexicon-based automatic labeling provides more consistent and reliable sentiment classification, particularly in specialized domains like video game adaptation reviews. Integrating VADER labeling with SVM enhances sentiment analysis effectiveness and offers practical value for media analytics, content creation, and audience insight research.
Sentiment Classification of MyPertamina Reviews Using Naïve Bayes and Logistic Regression Dwi Yuni Saraswati; Handayani, Maya Rini; Umam, Khothibul; Mustofa, Mokhamad Iklil
Journal of Applied Informatics and Computing Vol. 9 No. 4 (2025): August 2025
Publisher : Politeknik Negeri Batam

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.30871/jaic.v9i4.9723

Abstract

This research conducts a comparative evaluation of the effectiveness of the Naïve Bayes and Logistic Regression algorithms in mapping public perceptions of the MyPertamina application on the Google Play Store. The data consists of 2,000 user reviews obtained through a scraping technique. The research steps include labeling the reviews as positive or negative, followed by pre-processing and TF-IDF weighting. The dataset was systematically divided into two parts, with 80% allocated for model training and the remaining 20% for evaluation. The Naïve Bayes and Logistic Regression models were implemented using the Python programming language and evaluated based on accuracy, precision, recall, and F1-score metrics. The analysis shows that Logistic Regression achieved an accuracy of 86%, while Naïve Bayes achieved 81%. Logistic Regression demonstrated superior performance as it effectively captures linear relationships between features in TF-IDF representations and provides a more balanced outcome in terms of precision and recall. In contrast, Naïve Bayes is more influenced by high-frequency word distributions and does not account for feature correlations, which can limit its performance in certain contexts. Therefore, Logistic Regression is considered more suitable for sentiment classification tasks in this study. These findings emphasize the importance of selecting appropriate algorithms for sentiment analysis and suggest opportunities for future research using alternative methods to enhance predictive accuracy.
IKN Public Opinion on TikTok Before and After Efficiency Policy: CNN-LSTM on Imbalanced Data Sufiya, Ikhwanus; Umam, Khotibul; Handayani, Maya Rini
Jurnal Pendidikan Informatika (EDUMATIC) Vol 9 No 2 (2025): Edumatic: Jurnal Pendidikan Informatika
Publisher : Universitas Hamzanwadi

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.29408/edumatic.v9i2.30123

Abstract

Growing polarization in Ibu Kota Nusantara (IKN) stems from conventional sentiment analysis tools’ inability to decode TikTok’s contextual complexities, particularly multimodal sarcasm and vernacular-policy relationships (e.g., mangkrak for project cancellations). This study develops a policy-aware hybrid model (CNN-BiLSTM + Policy Knowledge Graph) to decode TikTok’s multimodal sarcasm and vernacular-policy links (e.g., mangkrak), enabling: youth sentiment quantification post-IKN’s 73.3% budget cuts, social criticism-socio-political reality mapping, and evidence-based interventions mitigating Global South strategic project polarization. Using the Knowledge Discovery in Databases framework, we analyzed 2,950 high-engagement TikTok comments (≥10 interactions) from verified accounts (@Polindo.id and @geraldvincentt) across two periods: pre-policy (June-August 2024) and post-policy (January-March 2025). Methodologically, slang normalization, stemming, and minority-class weighting (15×) preceded classification via a CNN-BiLSTM architecture integrated with Policy Knowledge Graphs. Results showed an 18.88% reduction in negative sentiment (83.2%-8.7%), model accuracy of 94.13% (AUC-PR 0.91), and strong correlations between vernacular terms (e.g., mandek [stagnation]) and policy outcomes (r = -0.89; p < 0.01), with investor asing mentions surging 463% post-policy. These validate deep learning-enabled social listening for real-time policy diagnostics, with implications for fiscal transparency dashboards, algorithmic bias mitigation, and context-driven policy communication prioritizing vulnerable groups in SDG infrastructure governance.
Implementasi Algoritma Random Forest dalam Klasifikasi Ulasan Pengunjung Mall Semarang untuk Pengambilan Keputusan Layanan Maizaliyanti, Annisa; Umam, Khothibul; Yuniarti, Wenty Dwi; Handayani, Maya Rini
Jurnal Pendidikan Informatika (EDUMATIC) Vol 9 No 2 (2025): Edumatic: Jurnal Pendidikan Informatika
Publisher : Universitas Hamzanwadi

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.29408/edumatic.v9i2.30379

Abstract

Visitor preferences for malls in Semarang are not optimal because bold reviews have not been utilized optimally in decision making. Our research aims to classify the sentiment of Google Maps reviews from 13 malls in Semarang with a total of 2,600 reviews. Labeling is done manually based on ratings, where ratings 1–3 are considered negative reviews and 4–5 as positive reviews. The classification method used is Random Forest because the ensemble approach (bagging) provides optimal results. The research process includes data collection, labeling, cleaning, data sharing, classification, and model evaluation. The data used is unbalanced and dominated by positive reviews, so the Synthetic Minority Over-sampling Technique (SMOTE) technique was applied. The overall accuracy before and after SMOTE remained the same at 84%. However, the model's performance in detecting negative reviews increased from 27% to 44% in recall and F1-score from 0.40 to 0.52, but these values ​​are still relatively low. Java Supermall Semarang is the mall with the best reviews, with a classification accuracy reaching 90%. This model is better at recognizing positive reviews, but less reliable for negative reviews. Therefore, its use as a decision-making preference needs to be done with caution. This research opens up opportunities for further development, including the use of other models such as BERT which are superior in understanding context and language in reviews.
Klasifikasi sentimen pada ulasan pengguna aplikasi Cryptocurrency di Google Play Store menggunakan algoritma Decision Tree Tsuroyya, Kamiliya; Umam, Khothibulu; Yuniarti, Wenty Dwi; Handayani, Maya Rini
AITI Vol 22 No 2 (2025)
Publisher : Fakultas Teknologi Informasi Universitas Kristen Satya Wacana

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.24246/aiti.v22i2.279-293

Abstract

Cryptocurrency has become a trend in digital investment. The Pintu application exemplifies the use of digital technology for trading cryptocurrency assets. Reviews from the Google Play Store serve as an important source of data to understand the opinions of Pintu application users. This study focuses on investigating the sentiment analysis of Pintu application users sourced from the Google Play Store by implementing the Decision Tree and Random Forest algorithms. The approach used involves collecting data from the Google Play Store, which contains user reviews and ratings. The data is then labeled as positive or negative and cleaned, processed, and analyzed using Decision Tree and Random Forest algorithms. The results of the study showed that the accuracy of the Decision Tree reached 0.90, while the Random Forest achieved an accuracy of 0.88. From these results, it can be concluded that the Decision Tree is superior in classifying text mining with high accuracy. The difference between the two methods is insignificant in terms of accuracy, specifically for Decision Tree, with an accuracy of 0.90, Precision of 0.91, and recall of 0.95, and Random Forest, with an accuracy of 0.88, precision of 0.87, and recall of 0.95. User sentiment analysis of the Pintu application provides a positive response to using the Pintu application.
Perbandingan Klasifikasi Single-Label dan Multi-Label Ulasan Pengguna Lapangan Futsal di Semarang Menggunakan SVM Syifa, Achrijal Shohib Arya; Umam, Khotibul; Handayani, Maya Rini; Aini, Siti Nur
InComTech : Jurnal Telekomunikasi dan Komputer Vol 15, No 2 (2025)
Publisher : Department of Electrical Engineering

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.22441/incomtech.v15i2.33905

Abstract

Futsal merupakan cabang olahraga yang semakin populer di seluruh Indonesia, termasuk di Semarang. Penelitian ini bertujuan untuk melakukan klasifikasi sentimen ulasan pengguna mengenai lapangan futsal di Kota Semarang menggunakan metode Support Vector Machine. Data penelitian diperoleh melalui scraping ulasan Google Maps dengan ekstensi Chrome “Instant Data Scraper” dan terdiri dari 1.189 ulasan. Proses penelitian mencakup pengumpulan data, Cleaning dan pre-processing (normalisasi teks, modifikasi data, tokenisasi, stop word filtering, stemming), pelabelan (single label dan multi label), pembagian data (80% pelatihan dan 20% pengujian), pemodelan menggunakan SVM (single label dengan GridSearchCV dan multi label dengan One-vs-Rest Classifier), serta evaluasi model dengan metrik presisi, recall, dan F1-Score. Hasil menunjukkan pemodelan Support Vector Machine single-label mencapai presisi 0,84, recall 0,73, dan F1-Score 0,78. Sementara pemodelan Support Vector Machine multi-label mencapai presisi 0,96, recall 0,88, dan F1-Score 0.92. Dari ulasan yang dinalisis, sebaran data pada single-label maupun multi-label menunjukan dominasi ulasan kategori Fasilitas, menegaskan bahwa Fasilitas merupakan kategori yang paling sering dikomentari oleh pengguna. Temuan ini tidak hanya memberikan wawasan praktis bagi pengelola lapangan futsal, tetapi juga berkontribusi pada pengembangan metode klasifikasi ulasan berbasis machine learning dalam domain analisis opini, khususnya dalam membandingkan performa pendekatan single-label dan multi-label pada data multi-kategori di bidang teknologi informasi.
Identification of Buzzers in Skincare Reviews Using a Lexicon-Based Sentiment Analysis Method Pramesti, Arfiana Diah; Umam, Khothibul; Handayani, Maya Rini
Journal of Applied Informatics and Computing Vol. 9 No. 5 (2025): October 2025
Publisher : Politeknik Negeri Batam

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.30871/jaic.v9i5.11005

Abstract

Along with the rapid development of digital technology, social media has become the main platform for consumers to share experiences about products, including skincare products. However, it is not uncommon for reviews provided by users to not reflect authentic experiences, but rather reviews created by certain parties, or buzzers, to manipulate public perception. The presence of buzzers in skincare reviews is important to consider, as they can affect consumer trust and influence purchasing decisions. This study aims to identify the presence of buzzers in skincare product reviews using a lexicon dictionary-based sentiment analysis. Of the 529 comments analyzed, 75 comments showed negative sentiment and 454 comments showed positive sentiment. The classification results revealed that 85.8% of the comments belonged to the non-buzzer category, while 14.2% were indicated as buzzers. Evaluation of the classification model showed high accuracy, reaching 93%, but performance in detecting buzzers was limited, with a recall metric of only 0.50. This shows that while the model managed to classify non-buzzer comments well, there are still difficulties in identifying buzzer comments, mostly due to data imbalance. This research emphasizes the importance of a proper analytical approach in detecting inauthentic reviews to ensure the information consumers receive remains accurate, transparent, and accountable.
Mapping the Polarity of Tourist Opinions on Indonesian Destinations through Google Maps Reviews Using Supervised Learning Methods Sa’adah, Siti Miftahus; Umam, Khothibul; Handayani, Maya Rini; Mustofa, Mokhammad Iklil
Journal of Applied Informatics and Computing Vol. 9 No. 5 (2025): October 2025
Publisher : Politeknik Negeri Batam

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.30871/jaic.v9i5.9836

Abstract

The advancement of information technology has transformed how individuals seek information and plan their travels, notably through online reviews of tourist attractions on platforms like Google Maps. However, these reviews do not always align with visitors' expectations, necessitating further analysis to comprehend the underlying sentiments. The objective of this research is to inspect the performance of multiple machine learning algorithms in executing sentiment analysis on user generated reviews related to tourist attractions in Indonesia. The algorithms examined include Multinomial Naïve Bayes, Random Forest Classifier, Logistic Regression, Support Vector Machine, K-Nearest Neighbors, and Extra Trees Classifier. The research process encompasses data collection and labeling, data preprocessing, exploratory data analysis (EDA), Word Cloud visualization, feature extraction, classification implementation, and performance evaluation. Experimental results indicate that the K-Nearest Neighbors (KNN) algorithm attain the most accuracy and F1-score of 97%, indicating its effectiveness in categorizing text-based sentiment reviews sourced from the Google Maps platform.
THE PERCEPTIONS OF SEMARANG FIVE STAR HOTEL TOURISTS WITH SUPPORT VECTOR MACHINE ON GOOGLE REVIEWS Aufan, Muhammad Haikal; Handayani, Maya Rini; Nurjanna, Afifah Basmah; Wibowo, Nur Cahyo Hendro; Umam, Khotibul
Jurnal Teknik Informatika (Jutif) Vol. 5 No. 5 (2024): JUTIF Volume 5, Number 5, Oktober 2024
Publisher : Informatika, Universitas Jenderal Soedirman

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.52436/1.jutif.2024.5.5.2025

Abstract

Travelers on the road sometimes need a hotel to rest. In choosing a hotel, they refer to the ratings or reviews written by users through reviews on Google. This is because not all star hotels provide facilities in accordance with user assessments. This study discusses the analysis of the opinions of tourists who have stayed in 5-star hotels in Semarang through a review of commentary data on Google. The 5-star hotels used as the research are Padma, Gumaya, Tentrem, Grand Candi, Ciputra, and PO. The dataset of the six hotels was obtained through a scraping process then followed by data pre-processing. The data was retrieved from Google Maps using the Chrome Instant Data Scrapper extension. Data preprocessing begins with case folding, tokenizing, filtering, and ends with stemming. Support Vector Machine (SVM) is implemented for sentimen classification process. The results from this study are the majority of 5-star hotel reviews in Semarang tend to have positive rather than negative sentimens. Our model was able to produce an accuracy of 0.87 to 0.98. The highest accuracy was achieved by Ciputra Hotel at 0.98 with 543 positive reviews.
Identifikasi Polaritas Sikap Pengguna Aplikasi X terhadap Coretax di Indonesia Menggunakan Algoritma Naïve Bayes Prasilda, Dina Rahma; Yuniarti, Wenty Dwi; Handayani, Maya Rini; Umam, Khothibul
JURNAL RISET KOMPUTER (JURIKOM) Vol. 12 No. 3 (2025): Juni 2025
Publisher : Universitas Budi Darma

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.30865/jurikom.v12i3.8548

Abstract

The Core Tax Administration System (Coretax) was launched by the Directorate General of Taxes (DGT) in January 2025 as a technology-based integrated tax system. While its initial goal was to improve tax efficiency and compliance, Coretax faced technical challenges, including system errors, slow processing speed, and criticism from the public. The main platform used to address these challenges is the X app (formerly known as Twitter). This research aims to understand the public's views and responses to Coretax's services by analyzing user sentiment patterns seen on social media. The research identifies the polarity of user attitudes by utilizing natural language processing (NLP) and Naïve Bayes algorithms, applied to a dataset of 1,628 tweets collected between January and March 2025. The analyzed data reflects a wide range of public reactions that include both positive and negative opinions towards the Coretax implementation, both in terms of functionality and ease of use. The results show that the model has an accuracy rate of 93.07%, a precision value of 95%, a recall value of 96%, and an F1-Score value of 96%. The results of this study are expected to be able to provide precise mapping related to changes in public opinion towards Coretax, so that it can be a valuable source of information for application developers, policy makers in the field of taxation, and analysis in the technology sector in responding to the needs and expectations of society in the digital era.
Co-Authors ., Kumarudin ., Kumarudin Amal, Muhammad Niltal Amelia Rahmi Apriliyani, Meli Arroyan, Devina Asep Dadang Abdullah Asep Dadang Abdullah Aufan, Muhammad Haikal Azziizah, Almira Farradinda Chairullah, Dimas Dina Wulan Yekti rahayu Dwi Yuni Saraswati Dwi Yuniarti, Wenty Ema Hidayanti Fastabiqul Khusna Febrianto, Bagus Fiashintha Dewi Hanya Abriananta Heti Aprilianti Hidayati, Ema Jinan, Muhammad Syifaaul Khoirotulmuadiba Purifyregalia Khoirul Adib Khothibul Umam Khothibul Umam Khothibul Umam Khotibul Umam Maizaliyanti, Annisa Malikhatul Ibriza Masuzzahra, Tsaura Rafah Masy Ari Ulinuha Mokhamad Iklil Mustofa Muhadzib Al-Faruq, Muhammad Naufal Muhammad Rafid Pratama Mustofa Hilmi Mustofa, Hery Mustofa, Mokhammad Iklil Musyafak, Najahan Musyafak, Najahan Musyaffaq, Mirza Izzal Nahdhudin, Muhammad Nikmal Maulana Nur Cahyo Hendro Wibowo Nur Cahyo Hendro Wibowo Nur Cahyo Hendro Wibowo, Nur Cahyo Hendro Nurjanna, Afifah Basmah Nur’aini, Siti Pramesti, Arfiana Diah Prasilda, Dina Rahma putri lathifah, shofi Putri, Indira Alifia Qonita, Nuurun Najmi Rahmadani, Nurul Robby Kurniawan Budhi Safitri, Sindy Eka Sajmira, Danita Divka Salmalina, Divana Taricha Saputra, Adika Kaka Sa’adah, Siti Miftahus Septyorini, Talitha Dwi Siti Hikmah Siti Hikmah, Siti Siti Nur Aini, Siti Nur Siti Nur’aini Sufiya, Ikhwanus Syifa, Achrijal Shohib Arya Tsuroyya, Kamiliya Umam, Khothibulu Vensy Vydia wening wihartati, wening Wenty Dwi Yuniarti Wenty Dwi Yuniarti Wenty Dwi Yuniarti Wenty Dwi Yuniarti, Wenty Dwi