Claim Missing Document
Check
Articles

Enhancing Software Defect Prediction: HHO-Based Wrapper Feature Selection with Ensemble Methods Fauzan Luthfi, Achmad; Herteno, Rudy; Abadi, Friska; Adi Nugroho, Radityo; Itqan Mazdadi, Muhammad; Athavale, Vijay Anant
Indonesian Journal of Electronics, Electromedical Engineering, and Medical Informatics Vol. 7 No. 2 (2025): May
Publisher : Jurusan Teknik Elektromedik, Politeknik Kesehatan Kemenkes Surabaya, Indonesia

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.35882/f2140043

Abstract

The growing complexity of data across domains highlights the need for effective classification models capable of addressing issues such as class imbalance and feature redundancy. The NASA MDP dataset poses such challenges due to its diverse characteristics and highly imbalanced classes, which can significantly affect model accuracy. This study proposes a robust classification framework integrating advanced preprocessing, optimization-based feature selection, and ensemble learning techniques to enhance predictive performance. The preprocessing phase involved z-score standardization and robust scaling to normalize data while reducing the impact of outliers. To address class imbalance, the ADASYN technique was employed. Feature selection was performed using Binary Harris Hawk Optimization (BHHO), with K-Nearest Neighbor (KNN) used as an evaluator to determine the most relevant features. Classification models including Random Forest (RF), Support Vector Machine (SVM), and Stacking were evaluated using performance metrics such as accuracy, AUC, precision, recall, and F1-measure. Experimental results indicated that the Stacking model achieved superior performance in several datasets, with the MC1 dataset yielding an accuracy of 0.998 and an AUC of 1.000. However, statistical significance testing revealed that not all observed improvements were meaningful; for example, Stacking significantly outperformed SVM but did not show a significant difference when compared to RF in terms of AUC. This underlines the importance of aligning model choice with dataset characteristics. In conclusion, the integration of advanced preprocessing and metaheuristic optimization contributes positively to software defect prediction. Future research should consider more diverse datasets, alternative optimization techniques, and explainable AI to further enhance model reliability and interpretability.
Multi-Criteria Decision Making dalam Seleksi Fitur Ensemble untuk Prediksi Cacat Perangkat Lunak Fikri, Muhammad; Herteno, Rudy; Adi Nugroho, Radityo; Wahyu Saputro, Setyo; Abadi, Friska
Jurnal Teknologi Informasi dan Ilmu Komputer Vol 12 No 6: Desember 2025
Publisher : Fakultas Ilmu Komputer, Universitas Brawijaya

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.25126/jtiik.2025125

Abstract

Prediksi cacat perangkat lunak merupakan upaya strategis dalam meningkatkan kualitas produk melalui identifikasi dini modul yang berpotensi cacat. Kinerja prediksi dipengaruhi oleh pemilihan fitur, karena informasi yang berlebihan dan tidak relevan dapat mempengaruhi kualitas pembelajaran model. Seleksi fitur ensemble dinilai efektif dalam menyeleksi fitur yang relevan dengan menggabungkan beberapa metode seleksi fitur berbasis filter. Diperlukan mekanisme integrasi untuk menyatukan hasil dari empat teknik filter—Mutual Information, Fisher Score, Uncertainty dan Relief. Penelitian ini membandingkan empat metode Multi‑Criteria Decision Making—TOPSIS, VIKOR, EDAS, dan WASPAS—yang bekerja dengan merangking nilai relevansi fitur hasil seleksi filter tersebut. Sepuluh fitur teratas dari tiap metode kemudian dievaluasi menggunakan model Random Forest dengan metrik AUC melalui K‑Fold cross‑validation. Dari 12 dataset NASA MDP yang diuji, TOPSIS menunjukkan kinerja paling konsisten dan terbaik dengan nilai rata-rata AUC sebesar 0,8038. Temuan ini menegaskan pentingnya pemilihan metode integrasi yang tepat dalam meningkatkan akurasi prediksi cacat perangkat lunak dan memberikan panduan bagi pengembangan model yang lebih efektif.   Abstract Software defect prediction is a strategic effort to improve product quality through early identification of potentially defective modules. Prediction performance is influenced by feature selection, because redundant and irrelevant information can affect the quality of model learning. Ensemble feature selection is considered effective in selecting relevant features by combining several filter-based feature selection methods. An integration mechanism is needed to unify the results of four filter techniques—Mutual Information, Fisher Score, Uncertainty and Relief. This study compares four Multi-Criteria Decision Making methods—TOPSIS, VIKOR, EDAS, and WASPAS—which work by ranking the relevance values ​​of the filter-selected features. The top ten features from each method are then evaluated using the Random Forest model with the AUC metric through K-Fold cross-validation. Of the 12 NASA MDP datasets tested, TOPSIS showed the most consistent and best performance with an average AUC value of 0.8038. These findings emphasize the importance of choosing the right integration method in improving the accuracy of software defect prediction and provide guidance for the development of more effective models.
Analisis Sentimen Ulasan Media Sosial UMKM Kuliner dengan Pendekatan Lexicon-Based dan Kosakata Khusus Setyo Wahyu Saputro; Friska Abadi; Radityo Adi Nugroho
Jurnal Informatika Polinema Vol. 12 No. 2 (2026): Vol. 12 No. 2 (2026)
Publisher : UPT P2M State Polytechnic of Malang

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.33795/jip.v12i2.9302

Abstract

UMKM kuliner di Kalimantan Selatan memanfaatkan media sosial sebagai sarana utama untuk mengetahui opini pelanggan, namun jumlah komentar yang sangat besar menyulitkan pelaku usaha untuk menelaahnya secara manual. Kondisi ini menegaskan perlunya pendekatan analisis sentimen yang mampu mengolah data ulasan secara efisien serta sesuai dengan karakteristik bahasa lokal. Penelitian ini bertujuan mengembangkan metode analisis sentimen berbasis lexicon yang diperkaya dengan kosakata domain-spesifik kuliner dan bahasa Banjar agar hasil klasifikasi lebih akurat dan kontekstual. Data penelitian diperoleh dari 3.500 komentar publik di Instagram dan TikTok. Tahap preprocessing mencakup case folding, pembersihan karakter khusus, tokenisasi, stopword removal, normalisasi, dan stemming. Selanjutnya, InSet Lexicon disempurnakan melalui penyuntikan kosakata baru serta penyesuaian bobot kata sesuai konteks kuliner lokal. Hasil analisis menunjukkan distribusi sentimen terdiri dari 2.050 komentar positif (58,57%), 934 komentar netral (26,69%), dan 516 komentar negatif (14,74%). Evaluasi menunjukkan peningkatan akurasi signifikan setelah perluasan lexicon, yaitu 93,49% untuk sentimen negatif, 94,64% untuk netral, dan 96,94% untuk positif, dibandingkan akurasi awal yang berkisar antara 51–73%. Temuan ini membuktikan bahwa pengayaan lexicon menggunakan kosakata lokal dan domain-spesifik secara substansial meningkatkan performa analisis sentimen. Pendekatan ini memberikan solusi praktis dan terjangkau bagi UMKM untuk memahami opini pelanggan secara lebih representatif, serta dapat dimanfaatkan dalam pengambilan keputusan strategis dan perbaikan kualitas layanan maupun promosi produk kuliner.
Empirical Performance of E2E Frameworks in React-Vue SPAs Using DIA Rezeki, Abdillah; Saputro, Setyo Wahyu; Saragih, Triando Hamonangan; Nugroho, Radityo Adi; Abadi, Friska
International Journal of Advances in Data and Information Systems Vol. 7 No. 1 (2026): April 2026 - International Journal of Advances in Data and Information Systems
Publisher : Indonesian Scientific Journal

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.59395/ijadis.v7i1.1528

Abstract

Modern web applications increasingly adopt Single-Page Application (SPA) architectures to enhance the user experience through client-side rendering and dynamic content loading. However, these characteristics introduce significant challenges for automated end-to-end (E2E) testing, including asynchronous DOM manipulation, complex state management, and timing synchronization issues. This study presents a comprehensive empirical comparison of three prominent E2E testing frameworks—Selenium WebDriver, Cypress, and Playwright—across React and Vue-based SPAs. Using a quantitative experimental approach, 25 standardized test cases were executed 15 times each across Chrome, Firefox, and Edge, for a total of 270 testing sessions. Performance evaluation focused on four key metrics: execution time, success rate, CPU usage, and memory consumption. Results demonstrate that Playwright achieved the fastest execution time (56.25 seconds on React-Chrome), while Selenium exhibited superior resource efficiency with the lowest memory consumption (196.59 MB on Vue-Chrome). The Distance to Ideal Alternative (DIA) multi-criteria decision analysis method identified Playwright-Chrome as optimal for React applications (DIA score: 0.886715) and Selenium-Chrome for Vue applications (DIA score: 0.908237), indicating that framework selection should be context-dependent based on application characteristics and deployment requirements. This research supports the conclusion that no universal "best" testing framework exists, underscoring the importance of evidence-based, application-specific tool selection in software quality assurance.
Metrics Based Feature Selection for Software Defect Prediction Radityo Adi Nugroho; Friska Abadi; M. Reza Faisal; Rudy Herteno; Rahmat Ramadhani
Jurnal Komputasi Vol. 8 No. 2 (2020)
Publisher : Jurusan Ilmu Komputer Fakultas MIPA Universitas Lampung

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.23960/komputasi.v8i2.2670

Abstract

Nowadays, software is very influential on various sectors of life, both to solve business needs, as well as personal needs. To have a Software with high quality, testing is needed to avoid software defect. Research on software defects involving Machine Learning is currently being carried out by many researchers. This method contains one important step, which is called feature selection. In this study, researchers conducted a feature selection based on the software metric category to determine the level of accuracy of the prediction of software defects by utilizing 13 (thirteen) datasets from NASA MDP namely CM1, JM1, KC1, KC3, KC4, MC1, MC2, MW1, PC1, PC2, PC3, PC4, and PC5. To classify, the researchers involved 5 (five) classifiers, namely Naive Bayes, Decision Trees, Random Forests, K-Nearest Neighbor, and Support Vector Machines. The research result shows that each attribure on software metric categories has effect on each dataset. Naive Bayes Algorithm and Random Forest Algorithm can give better performance than other algorithm in classifieng software defect with feature selection based on metrics. On the other hand, the best metrics category on each classifier algorithm is metric Misc. From average AUC value, it can be concluded that metrics category which can give best performance is metric LoC, followed by metric Misc. Both categories have achieved highest AUC value in Random Forest classifier.
Analisis Komparasi Implementasi Steganografi White-Space dan White-Space Modified pada Artikel Terenkripsi AES dalam HTML5 Rudy Herteno; Dodon Turianto Nugrahadi; Muhammad Sholih Afif; M Reza Faisal; Friska Abadi
Jurnal Komputasi Vol. 8 No. 1 (2020)
Publisher : Jurusan Ilmu Komputer Fakultas MIPA Universitas Lampung

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.23960/komputasi.v8i1.2525

Abstract

The level of internet usage continues to increase until now.  information exchange requires security that cannot be predicted by others.  one technique for securing information is steganography.  Steganography techniques are the science and art of hiding information.  This technique can hide the content of information in media that cannot be guessed by ordinary people, so as not to arouse suspicion of the people who see it.  One of the media that can implement the white-space modified steganography method is HTML pages.  in addition, AES (Advanced Encryption Standard) is a lighter encryption security algorithm compared to other algorithms. In this study, plain text that has been encrypted into cipher text is then inserted with white-space and white-space modification steganography techniques. Data changes have occurred but only less than 1 percent.  In experiments that have been implemented on Google Chrome and Mozilla Firefox are the same except in Internet Explorer, which changes the data slightly larger.The implementation of AES encryption and stegano white-space original, has 100% success but the 80% decryption process is successful, but the decryption results contain additional binaries. This happen because the use of tabulation (tabs) instead of spaces in HTML5 articles, and this is often found in HTML articles. while the implementation of AES encryption and stegano whitespace modified, has a success of 100% and the decryption process of 90% succeeded without any changes. 1 article failed because the number of articles is too small compared to the amount of space provided. The conclusion that implementation of AES encryption and white-space modified is more appropriate to be implemented in HTML5 articles, and than the use of tabulation and the number of characters also consequences on the implementation.Keywords: Information, Steganography, White-space modified, Security, AES, Web Browser 
Implementation of PPCA Imputation, SMOTE-N Class Balancing in Hepatitis Classification Using Naïve Bayes Siti Fathmah; Dwi Kartini; Friska Abadi; Irwan Budiman; Muhammad Itqan Mazdadi
JUITA: Jurnal Informatika JUITA Vol. 12 No. 2, November 2024
Publisher : Department of Informatics Engineering, Universitas Muhammadiyah Purwokerto

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.30595/juita.v12i2.21528

Abstract

The availability of complete data in research is crucial, especially in the initial stages. The Hepatitis data used in this study encountered issues such as missing data and class imbalance, which hindered its optimal utilization. The method employed to address missing data was the PPCA imputation method. After filling in the missing data, the data was balanced using the SMOTE-N class balancing method and classified using Gaussian Naïve Bayes. The aim of this research was to compare the classification evaluation of hepatitis disease using Naive Bayes with the PPCA imputation approach and SMOTE-N class balancing. The best results from each scenario yielded an AUC value of 0.833 in the first scenario with an 80:20 data split for training and testing, and 0.875 in the second scenario with a 90:10 data split. The highest AUC value was obtained in the application of PPCA imputation with SMOTE-N class balancing using Naive Bayes classification. This demonstrates that the implementation of PPCA imputation with SMOTE-N class balancing has a better impact on the performance of Naïve Bayes classification.
Quantifying the Impact of Text Preprocessing on IndoBERT Fine-Tuning for Indonesian Informal Culinary Sentiment Analysis Rahmat Budianoor; Setyo Wahyu Saputro; Friska Abadi; Radityo Adi Nugroho; Andi Farmadi
Journal of Computing Theories and Applications Vol. 3 No. 4 (2026): JCTA 3(4) 2026
Publisher : Universitas Dian Nuswantoro

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.62411/jcta.15980

Abstract

Indonesian culinary comments on social media platforms such as Instagram are characterized by informal spelling, regional language mixing, slang expressions, and emojis, posing substantial challenges for automated sentiment classification. While IndoBERT has demonstrated strong performance across Indonesian natural language processing tasks, the contribution of individual preprocessing components to fine-tuning performance on informal text remains underexplored, particularly in the culinary domain. This study addresses this gap by conducting a systematic preprocessing ablation study on IndoBERT-Base fine-tuning for Indonesian culinary sentiment classification, accompanied by a comparative evaluation against Naive Bayes with TF-IDF, SVM with TF-IDF, and BiLSTM as representative baselines. A dataset of 3,500 manually labeled Instagram culinary comments across three sentiment classes was used, with a stratified 80/10/10 split. Six preprocessing variants were evaluated under identical experimental conditions to isolate the contribution of each component. The results show that slang normalization is the most impactful single preprocessing step, yielding a macro F1-score gain of +0.0609 over the no-preprocessing baseline, while the full pipeline achieves an accuracy of 0.8800 and a macro F1-score of 0.8465. IndoBERT-Base with the full pipeline outperforms all baselines across all evaluation metrics. Per-class analysis reveals that the negative class achieves the lowest F1-score of 0.7600, with sarcastic expressions and Banjar regional vocabulary identified as primary sources of misclassification. These findings indicate that preprocessing decisions have a measurable and non-uniform effect on IndoBERT fine-tuning performance. In this study, slang normalization provides the most substantial individual contribution in bridging the vocabulary gap between informal user-generated text and the model’s pre-training distribution.
Evaluating CNN Robustness for Face Mask Classification under Environmental Variations Bagaskara Ridho Vandio; Fatma Indriani; Andi Farmadi; Dodon Turianto Nugrahadi; Friska Abadi
Journal of Embedded Systems, Security and Intelligent Systems Vol 7 No 2 (2026): June 2026
Publisher : Program Studi Teknik Komputer

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.59562/jessi.v7i2.2617

Abstract

Purpose – This study aims to analyze and compare the performance of ResNet50 and MobileNetV3 for multi-class face mask classification under various environmental conditions. Design/methods/approach – ResNet50 and MobileNetV3 are trained using transfer learning for three-class face mask classification and evaluated under normal conditions and environmental variations, including illumination changes, blur, low compression, and rotation. Findings – Experimental results show that ResNet50 achieves an accuracy of 94.32% under normal conditions, slightly outperforming MobileNetV3 at 94.10%. Under environmental variations, the largest performance degradation is observed under darkening and blur conditions, while low compression and rotation have relatively minor effects. ResNet50 demonstrates higher robustness across most perturbation settings, whereas MobileNetV3 provides competitive performance with substantially better computational efficiency. Research implications/limitations – This study is limited to a controlled evaluation using synthetic environmental perturbations on a single dataset and does not consider broader dataset diversity. Therefore, the findings should be interpreted within the evaluated experimental conditions. Originality/value – This study provides a comparative analysis of model robustness under controlled environmental perturbations, highlighting the trade-off between robustness and computational efficiency for face mask classification systems.
Characteristics ransomware stop/djvu remk and erqw variants with static-dinamic analysis Dodon Turianto Nugrahadi; Friska Abadi; Rudy Herteno; Muliadi Muliadi; Muhammad Alkaff; Muhammad Alvin Alfando
Computer Science and Information Technologies Vol 6, No 3: November 2025
Publisher : Institute of Advanced Engineering and Science

Show Abstract | Download Original | Original Source | Check in Google Scholar | DOI: 10.11591/csit.v6i3.p283-293

Abstract

Ransomware has developed into various new variants every year. One type of ransomware is STOP/DJVU, containing more than 240+ variants. This research to determine changes in differences characteristics and impact between ransomware variants STOP/DJVU remk, which is a variant from 2020, and the erqw variant from 2023, through a mixed-method research approach. Observation, simulation using mixing static and dynamic malware analysis methods. Both variants are from the Malware Bazaar site. The total characteristics based on dynamic analysis, the remk variant has 177, and the erqw variant has 190, which increased by 1.8%. The total characteristics based on static analysis, the remk variants have 586, and the erqw variants have 736, which increased by 5.7%. All characteristics from remk to erqw increasing in dynamic analysis, except the number of payloads that decreased about 20%. In static analysis, all characteristics from remk to erqw increase except the number of sections decreased about 1.5%. It can be the affected CPU performance, because the remk variant affects performance by increasing CPU work by 3.74%, while the erqw variant affects performance by reducing CPU work by 1.18%, both compared with normal CPU. which will affect the ransomware's destructive work and require changes in its handling.
Co-Authors A.A. Ketut Agung Cahyawan W AA Sudharmawan, AA Abdullayev, Vugar Achmad Zainudin Nur Adam Mukharil Bachtiar Adi Mu'Ammar, Rifqi Adinda Ayu Puspita Ramadhani Ahmad Juhdi Amalia, Raisa Andi Farmadi Andi Farmadi Andi Farmandi Arif, Nuuruddin Hamid Athavale, Vijay Anant Bagaskara Ridho Vandio budiman, irwan Deni Kurnia Dodon Turianto Nugrahadi Dwi Kartini Dwi Kartini, Dwi Emma Andini Fatma Indriani Fauzan Luthfi, Achmad Febrian, Muhamad Michael Halimah Hanafi, Muhammad Bashir Hariyady Hariyady Herteno, Rudy Indriani, Fatma Irwan Budiman Irwan Budiman Irwan Budiman Itqan Mazdadi, Muhammad Kartika, Najla Putri Khusnul Rahmi Maulidha M Kevin Warendra Mafazy, Muhammad Meftah Martalisa, Asri Mbeledogu, Njideka Nkemdilim Mera Kartika Delimayanti Muhamad Fawwaz Akbar Muhammad Alkaff Muhammad Alkaff Muhammad Alvin Alfando Muhammad Azmi Adhani Muhammad Denny Ersyadi Rahman Muhammad Fikri Muhammad Haekal Muhammad Itqan Mazdadi Muhammad Khairin Nahwan Muhammad Mirza Hafiz Yudianto Muhammad Nabil Muyassar Rahman Muhammad Nazar Gunawan Muhammad Noor Muhammad Reza Faisal, Muhammad Reza Muhammad Rizky Aulia Ramadhan Muhammad Sholih Afif Muliadi Muliadi Muliadi Aziz Muliadi Muliadi Muliadi Muliadi Nabella, Putri Nor Indrani Nugrahadi, Dodon Nurlatifah Amini Nursyifa Azizah Prastya, Septyan Eka Pratama, Muhammad Yoga Adha Puput Dani Prasetyo Adi Putri Nabella Raditya, Virgi Atha Radityo Adi Nugroho Rahman Hadi Rahman Rahmat Budianoor Rahmat Ramadhani Rahmina Ulfah Aflaha Reina Alya Rahma Reza Faisal, Mohammad Rezeki, Abdillah Rinaldi Riza Susanto Banner Rizal, Muhammad Nur Rizky Ananda, Muhammad Rizky, Muhammad Hevny Rudy Herteno Rudy Herteno Rudy Herteno SALLY LUTFIANI Saragih, Triando Hamonangan Sarah Monika Nooralifa Sa’diah, Halimatus Septyan Eka Prastya Setyo Wahyu Saputro Siti Fathmah Siti Napi'ah Syarif Maulana, Syarif Tri Mulyani Ulya, Azizatul Umar Ali Ahmad Vina Maulida, Vina Wahyu Dwi Styadi Yasmin Dwi Safitri Yunida, Rahmi