Claim Missing Document
Check
Articles

Machine Learning Implementation for Sentiment Analysis on X/Twitter: Case Study of Class Of Champions Event in Indonesia Hafizah, Rini; Saragih, Triando Hamonangan; Muliadi, Muliadi; Indriani, Fatma; Mazdadi, Muhammad Itqan
Indonesian Journal of Electronics, Electromedical Engineering, and Medical Informatics Vol. 7 No. 2 (2025): May
Publisher : Jurusan Teknik Elektromedik, Politeknik Kesehatan Kemenkes Surabaya, Indonesia

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.35882/ijeeemi.v7i2.81

Abstract

Sentiment analysis on social media is becoming an important approach in understanding public opinion towards an event. Twitter, as a microblogging platform, generates a large amount of data that can be utilized for this analysis. This study aims to evaluate and compare the performance of three classification algorithms, namely Support Vector Machine (SVM), Random Forest, and Extreme Gradient Boosting (XGBoost), in sentiment analysis related to the Clash of Champions event in Indonesia. To represent the text data, two feature extraction techniques are used, namely Term Frequency-Inverse Document Frequency (TF-IDF) and Bag of Words (BoW). In addition, Synthetic Minority Over-sampling Technique (SMOTE) is applied to handle data imbalance, while model optimization is performed using GridSearchCV. The research dataset consists of 1,000 tweets collected through web scraping, then manually processed and labeled before model training and testing. The results showed that the TF-IDF technique provided superior results compared to BoW. The Random Forest model with TF-IDF achieved the highest accuracy of 91%, while XGBoost with TF-IDF had the highest Area Under the Curve (AUC) of 0.91. The findings confirm that the selection of appropriate feature extraction techniques and algorithms can improve accuracy in sentiment analysis. This study can be applied in public opinion monitoring and data-driven decision-making. Future research can explore word embedding techniques and transformer-based deep learning models to improve semantic understanding and accuracy of sentiment analysis.
Application of Adaboost Algorithm with SMOTE and Optuna Techniques in Sleep Disorder Classification Anshory, Muhammad Naufal; Mazdadi, Muhammad Itqan; Saragih, Triando Hamonangan; Budiman, Irwan; Saputro, Setyo Wahyu
Indonesian Journal of Electronics, Electromedical Engineering, and Medical Informatics Vol. 7 No. 2 (2025): May
Publisher : Jurusan Teknik Elektromedik, Politeknik Kesehatan Kemenkes Surabaya, Indonesia

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.35882/ijeeemi.v7i2.99

Abstract

Data imbalance is a serious challenge in developing machine learning models for sleep disorder classification. When models are trained on an uneven distribution of classes, classification performance for minority classes such as insomnia and sleep apnea is often low. As a result, the overall accuracy may seem elevated, yet the sensitivity to important cases to be weak. Therefore, this research aims to design and develop a robust sleep disorder classification model with the AdaBoost algorithm, with improved performance through the integration of two main approaches, namely data balancing technique utilizing SMOTE and hyperparameter optimization using Optuna. This research contributes by showing that the combination of the two approaches can significantly improve model performance, not only in terms of global accuracy, but also accuracy on previously overlooked minority classes. The dataset utilized is the Sleep Health and Lifestyle Dataset which consists of 374 synthesized data and is divided into three categories: insomnia, sleep apnea, and none. This method stages include data preprocessing, data division using train-test split (80:20), application of SMOTE to balance the class distribution, hyperparameter tuning using Optuna, and model training with the AdaBoost algorithm. Evaluation was performed using classification metrics: accuracy, precision, recall, and F1-score. Results showed that mix of SMOTE and Optuna yielded the best results, accuracy 90.6%, F1-score 0.83871 for insomnia, and 0.81250 for sleep apnea. This performance was consistently superior to scenarios with no SMOTE or no tuning. This confirms the importance of using combination strategies to obtain fair and accurate classification on medical data. Future research is recommended to use real datasets as well as test the capabilities of this research on other models such as XGBoost or LightGBM.
Revitalisasi Pengemasan Produk UMKM “Woro Production” sebagai Upaya Peningkatan Daya Saing Melalui Penerapan Teknologi Inovatif Mazdadi, Muhammad Itqan; Sari, Anna Khumaira; Normaidah, Normaidah; Saputra, Adryan Maulana; Rahmah, Indah Noor; Ramadhani, Muhammad Irfan; Rahmawati, Nanda Hesti
Jurnal Pengabdian UNDIKMA Vol. 6 No. 4 (2025): November
Publisher : LPPM Universitas Pendidikan Mandalika (UNDIKMA)

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.33394/jpu.v6i4.17645

Abstract

This community service program aims to strengthen the capacity and technical skills of the “Woro Production” MSME by providing modern packaging equipment and training on its use to improve product efficiency and competitiveness. The implementation method involved training sessions and packaging simulations. Evaluation instruments included observation sheets and interviews to assess the partner’s skills in operating the packaging machine, and the resulting data were analyzed descriptively. The outcomes of this program indicate that participants were able to operate the equipment effectively, and the packaged products demonstrated improved hygiene, practicality, and visual appeal. This initiative is expected to enhance the competitiveness of Woro Production in local, national, and global markets.
KNN-MVO-SMOTE Algorithm for Air Quality Imbalanced Data Classification Rizky, Muhammad Miftahur; Mazdadi, Muhammad Itqan; Muliadi, Muliadi; Faisal, Mohammad Reza; Indriani, Fatma; Rozaq, Hasri Akbar Awal; Yildiz, Oktay
International Journal of Advances in Data and Information Systems Vol. 6 No. 3 (2025): December 2025 - International Journal of Advances in Data and Information Syste
Publisher : Indonesian Scientific Journal

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.59395/ijadis.v6i3.1424

Abstract

This research addresses air pollution, a pressing global issue influenced by geographic and temporal factors, using advanced machine-learning techniques to enhance air quality classification. By integrating the K-Nearest Neighbors (KNN) algorithm with the Synthetic Minority Over-sampling Technique (SMOTE) and Multi-Verse Optimization (MVO), we tackle challenges like data imbalance and parameter optimization. Our novel approach, which combines SMOTE and MVO within the KNN framework, has significantly increased classification accuracy to 97%, substantially improving over previous methods. The dataset includes diverse geographic and temporal data, with potential biases acknowledged and addressed. This study highlights the efficacy of merging MVO and SMOTE to optimize classification models, making a substantial contribution to environmental analysis and the fight against air pollution. Future research will explore AutoML technology to improve algorithmic optimization, offering more efficient and adaptive solutions. This pioneering effort emphasizes the critical role of technological innovation in tackling environmental challenges and marks a significant advancement in combating global air pollution.
Peningkatan Akurasi Model Boosting pada Prediksi Kesehatan Tidur Menggunakan Optuna Mazdadi, Muhammad Itqan; Saragih, Triando Hamonangan; Budiman, Irwan; Anshory, Muhammad Naufal
Jurnal Informatika Polinema Vol. 12 No. 2 (2026): Vol. 12 No. 2 (2026)
Publisher : UPT P2M State Polytechnic of Malang

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.33795/jip.v12i2.8878

Abstract

Kualitas tidur memiliki peran penting dalam menjaga kesehatan fisik maupun mental, sementara gangguan tidur dapat meningkatkan risiko berbagai penyakit kronis. Perkembangan machine learning membuka peluang untuk melakukan prediksi kesehatan tidur secara lebih akurat melalui pemanfaatan data gaya hidup. Penelitian ini berfokus pada penerapan algoritma boosting, yaitu XGBoost, LightGBM, AdaBoost, dan GradientBoosting, dengan dukungan teknik hyperparameter tuning berbasis Optuna untuk meningkatkan akurasi prediksi. Dataset yang digunakan adalah Sleep Health and Lifestyle Dataset yang memuat variabel demografis, kebiasaan hidup, serta kondisi tidur. Tahapan penelitian meliputi praproses data, pembagian data latih dan uji, pelatihan model, optimasi hyperparameter menggunakan Optuna dengan metode Tree-structured Parzen Estimator (TPE), serta evaluasi model menggunakan metrik akurasi. Hasil eksperimen menunjukkan bahwa tuning dengan Optuna memberikan peningkatan akurasi pada beberapa model, khususnya LightGBM dan AdaBoost, dengan nilai akurasi mencapai 93,3% dan 90,7%. Sementara itu, XGBoost dan GradientBoosting menunjukkan performa stabil dengan akurasi tetap tinggi baik sebelum maupun sesudah tuning. Temuan ini menegaskan bahwa efektivitas tuning bergantung pada karakteristik algoritma yang digunakan. Secara keseluruhan, penelitian ini membuktikan bahwa Optuna dapat menjadi solusi efektif dalam meningkatkan kinerja model boosting untuk prediksi kesehatan tidur. Sebagai arah penelitian lanjutan, disarankan penggunaan metrik evaluasi yang lebih beragam, penerapan teknik penyeimbangan data, serta eksplorasi integrasi dengan metode deep learning untuk memperkaya hasil analisis.
Comparasion Of Weather Classification Methods On Weather Images Using GLCM Features With Random Forest And Catboost Algoritms Noorhafizi, Muhammad; Saragih, Triando Hamonangan; Mazdadi, Muhammad Itqan; Muliadi, Muliadi; Herteno, Rudy; Rozaq, Hasri Awal Akbar
International Journal of Advances in Data and Information Systems Vol. 7 No. 1 (2026): April 2026 - International Journal of Advances in Data and Information Systems
Publisher : Indonesian Scientific Journal

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.59395/ijadis.v7i1.1456

Abstract

Weather image classification is an essential process for improving automated weather information systems. However, most existing studies rely on numerical meteorological data and rarely utilize the textural characteristics embedded in atmospheric imagery. This study addresses that limitation by applying the Gray Level Co-Occurrence Matrix (GLCM) for texture feature extraction combined with Random Forest (RF) and CatBoost algorithms for classification. The dataset, obtained from Kaggle, consists of 1,125 weather images categorized into four classes: cloudy, rain, shine, and sunrise. All images were uniformly normalized and augmented using four rotation angles (0°, 45°, 90°, 135°). GLCM features were extracted with a pixel distance of 1 and gray-level quantization of 8, generating four statistical attributes: contrast, correlation, energy, and homogeneity. Both algorithms were optimized through parameter tuning and evaluated using a 5-fold cross-validation scheme with an 80:20 split ratio. Results show that the Random Forest model (n_estimators = 100, max_depth = 10, random_state = 42) achieved the highest accuracy of 92.43% (±1.12), precision of 92.50%, recall of 92.43%, and F1-score of 92.42%. In comparison, CatBoost (iterations = 100, learning_rate = 0.1, depth = 6) achieved an accuracy of 68.88% (±2.31). The findings demonstrate that GLCM feature extraction combined with Random Forest offers superior stability and accuracy for weather image classification, providing a foundation for efficient and interpretable weather information systems.
Klasifikasi Tanaman Jarak Pagar Menggunakan Algoritme Deep Learning H2O Muhammad Itqan Mazdadi; Rahmat Ramadhani; Triando Hamonangan Saragih; Muhammad Haekal
Jurnal Komputasi Vol. 9 No. 1 (2021)
Publisher : Jurusan Ilmu Komputer Fakultas MIPA Universitas Lampung

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.23960/komputasi.v9i1.2774

Abstract

Tanaman jarak pagar merupakan tanaman multi fungsi yang memiliki banyak manfaat dari daun hingga buah. Tanaman jarak pagar sering digunakan untuk produk kecantikan hingga pengganti biodiesel. Penyakit yang menyerang tanaman jarak pagar dapat mengganggu hasil dari tanaman jarak pagar. Kurangnya pakar dibidang ini dan pengetahuan yang dimiliki petani menyebabkan sesuatu yang buruk. Persoalan ini dapat diselesaikan dengan metode Deep Learning. Metode Deep Learning yang digunakan adalah H2O. H2O digunakan karena dapat memberikan hasil komputasi yang cepat dan bisa memberikan akurasi yang baik. Pada penelitian ini bisa kita lihat bahwa H2O memberikan akurasi rata-rata maksimal sebesar 96,066% dengan parameter uji kombinasi data latih dan data uji 60:40, menggunakan satu layer dan jumlah epoch sebanyak 100. Pada penelitian ini membuktikan bahwa H2O bisa digunakan untuk identifikasi penyakit tanaman jarak pagar.
Comparison Algorithm for Diabetes Classification with Consideration of Mutual Information and Information Feature Rahmat Ramadhani; Triando Hamonangan Saragih; Muhammad Itqan Mazdadi; Muliadi Muliadi
Jurnal Komputasi Vol. 11 No. 1 (2023)
Publisher : Jurusan Ilmu Komputer Fakultas MIPA Universitas Lampung

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.23960/komputasi.v11i1.6649

Abstract

Diabetes is a prevalent disease in humans that is caused by excessive sugar levels in the body. If left untreated, it can lead to severe consequences such as paralysis, decay in certain parts of the body, and even death. Unfortunately, early detection of diabetes is difficult, and many cases go untreated until it is too late. However, the development of technology has opened up new possibilities for early detection and treatment of diabetes. One such approach is classification, a commonly used method in the field of Computer Science. Classification is used in various fields, including health, agriculture, and animal diseases, to draw conclusions based on input data using cause-and-effect relationships. Many different learning concepts and methods can be used in classification, with the Decision Tree concept being one of the most popular examples. This study compares several classification methods, including Decision Tree, Random Forest, AdaBoost, and Stochastic Gradient Boost, with feature selections carried out using MI and IF. The study aims to evaluate the effectiveness of these methods and the influence of feature selection on improving their performance. Based on the results of the study, it can be concluded that feature selection using Mutual Information and Importance Feature can improve the classification accuracy in some methods, particularly in Random Forest, AdaBoost, and Stochastic Gradient Boost. However, the Decision Tree algorithm did not show any improvement in accuracy after feature selection. The best classification accuracy was achieved with the Stochastic Gradient Boost method using the original dataset without feature selection, while the Random Forest method showed the highest accuracy after using all the features. Overall, the results suggest that feature selection can be a useful technique for improving the performance of classification algorithms in diabetes prediction. The study suggests that future research could investigate other classification methods, such as Neural Network or Deep Learning, and use optimization algorithms like Genetic Algorithm or Particle Swarm Optimization to improve feature selection results.
Implementation of PPCA Imputation, SMOTE-N Class Balancing in Hepatitis Classification Using Naïve Bayes Siti Fathmah; Dwi Kartini; Friska Abadi; Irwan Budiman; Muhammad Itqan Mazdadi
JUITA: Jurnal Informatika JUITA Vol. 12 No. 2, November 2024
Publisher : Department of Informatics Engineering, Universitas Muhammadiyah Purwokerto

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.30595/juita.v12i2.21528

Abstract

The availability of complete data in research is crucial, especially in the initial stages. The Hepatitis data used in this study encountered issues such as missing data and class imbalance, which hindered its optimal utilization. The method employed to address missing data was the PPCA imputation method. After filling in the missing data, the data was balanced using the SMOTE-N class balancing method and classified using Gaussian Naïve Bayes. The aim of this research was to compare the classification evaluation of hepatitis disease using Naive Bayes with the PPCA imputation approach and SMOTE-N class balancing. The best results from each scenario yielded an AUC value of 0.833 in the first scenario with an 80:20 data split for training and testing, and 0.875 in the second scenario with a 90:10 data split. The highest AUC value was obtained in the application of PPCA imputation with SMOTE-N class balancing using Naive Bayes classification. This demonstrates that the implementation of PPCA imputation with SMOTE-N class balancing has a better impact on the performance of Naïve Bayes classification.
Application of SMOTE to Handle Imbalance Class in Deposit Classification Using the Extreme Gradient Boosting Algorithm Dina Arifah; Triando Hamonangan Saragih; Dwi Kartini; Muliadi Muliadi; Muhammad Itqan Mazdadi
Jurnal Ilmiah Teknik Elektro Komputer dan Informatika Vol. 9 No. 2 (2023): June
Publisher : Universitas Ahmad Dahlan

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.26555/jiteki.v9i2.26155

Abstract

Deposits became one of the main products and funding sources for banks and increasing deposit marketing is very important. However, telemarketing as a form of deposit marketing is less effective and efficient as it requires calling every customer for deposit offers. Therefore, the identification of potential deposit customers was necessary so that telemarketing became more effective and efficient by targeting the right customers, thus improving bank marketing performance with the ultimate goal of increasing sources of funding for banks. To identify customers, data mining is used with the UCI Bank Marketing Dataset from a Portuguese banking institution. This dataset consists of 45,211 records with 17 attributes. The classification algorithm used is Extreme Gradient Boosting (XGBoost) which is suitable for large data. The data used has a high-class imbalance, with "yes" and "no" percentages of 11.7% and 88.3%, respectively. Therefore, the proposed solution in the research, which focused on addressing the Imbalance Class in the Bank marketing dataset, was to use Synthetic Minority Over-sampling (SMOTE) and the XGBoost method. The result of the XGBoost study was an accuracy of 0.91016, precision of 0.79476, recall of 0.72928, F1-Score of 0.56198, ROC Area of 0.93831, and AUCPR of 0.63886. After SMOTE was applied, the accuracy was 0.91072, the precision was 0.78883, the recall was 0.75588, F1-Score was 0.59153, ROC Area was 0.93723, and AUCPR was 0.63733. The results showed that XGBoost and SMOTE could outperform other algorithms such as K-Nearest Neighbor, Random Forest, Logistic Regression, Artificial Neural Network, Naïve Bayes, and Support Vector Machine in terms of accuracy. This study contributes to the development of effective machine learning models that can be used as a support system for information technology experts in the finance and banking industries to identify potential customers interested in subscribing to deposits and increasing bank funding sources.
Co-Authors AA Sudharmawan, AA Abdilah, Muhammad Fariz Fata Abdullayev, Vugar Ade Agung Harnawan, Ade Agung Adela Putri Ariyanti Afifa, Ridha Ahdyani, Annisa Salsabila Ahmad Rusadi Ahmad Rusadi Ahmad Rusadi Arrahimi - Universitas Lambung Mangkurat) Ahmad Rusadi Arrahimi - Universitas Lambung Mangkurat) Ahmad Shofi Khairian Ahmad Tajali Ahmad Tajali Aidil Akbar Al Ghifari, Muhammad Akmal Alamudin, Muhammad Faiq Amalia, Raisa Andi - Farmadi Andi Farmadi Andi Farmadi Andi Farmadi Anna Khumaira Sari Anshory, Muhammad Naufal Ansyari, Muhammad Ridho Antoh, Soterio Ardiansyah Sukma Wijaya Athavale, Vijay Anant Athavale, Vijay Annant budiman, irwan Buih, Putri Helena Junjung Deni Sutaji Dina Arifah Djordi Hadibaya Dodon Turianto Nugrahadi Dwi Kartini Dwi Kartini Dwi Kartini, Dwi Dzira Naufia Jawza Elvina Nur Hana Erdi, Muhammad Fatma Indriani Fatma Indriani Fitriani, Karlina Elreine Fitrinadi Friska Abadi Haekal, Muhammad Hafizah, Rini Helma Herlinda Herteno, Rudi Herteno, Rudy Indriani, Fatma Irwan Budiman Irwan Budiman Irwan Budiman Irwan Budiman Irwan Budiman Kenji Satou M. Apriannur M. Khairul Rezki Mafazy, Muhammad Meftah Muflih Ihza Rifatama Muhamad Fawwaz Akbar Muhamad Ihsanul Qamil Muhammad Haekal Muhammad Khairin Nahwan Muhammad Mada Muhammad Mirza Hafiz Yudianto Muhammad Mursyidan Amini Muhammad Reza Faisal, Muhammad Reza Muliadi Muliadi Muliadi Muliadi Muliadi Muliadi Muliadi Muliadi Muliadi Muliadi Muliadi Muliadi Nabella, Putri Noorhafizi, Muhammad Normaidah, Normaidah Nugraha, Muhammad Amir Nursyifa Azizah P., Chandrasekaran Patrick Ringkuangan Prastya, Septyan Eka Putri Nabella Radityo Adi Nugroho Rahmah, Indah Noor Rahmat Hidayat Rahmat Ramadhani Rahmat Ramadhani Rahmawati, Nanda Hesti Ramadhani, Muhammad Irfan Ramadhani, Rahmat Ratnapuri, Prima Happy Riadi, Agus Teguh Rifki Izdihar Oktvian Abas Pullah Rifki Rinaldi Rizky, Muhammad Miftahur Rozaq, Hasri Akbar Awal Rozaq, Hasri Awal Akbar Rudy Herteno Saputra, Adryan Maulana Saragih, Triando Hamonangan Satrio Yudho Prakoso Setyo Wahyu Saputro Shalehah Siti Fathmah Syahputra, Muhammad Reza Totok Wianto Wahyu Dwi Styadi Wijaya Kusuma, Arizha Yanche Kurniawan Mangalik YILDIZ, Oktay Yoga Pambudi Yudha Sulistiyo Wibowo Zaini Abdan