Indonesian Journal of Statistics and Its Applications
Vol 10 No 1 (2026): Vol 10 Issue 1 June 2026

A Two-Stage Framework for Unsupervised Sentiment Analysis with Model Selection and Semantic Similarity Evaluation

Cici Suhaeni (School of Data Science, Mathematics, and Informatics, IPB University)
Fani Fahira (School of Data Science, Mathematics, and Informatics, IPB University)
Hari Wijayanto (School of Data Science, Mathematics, and Informatics, IPB University)
La Ode Abdul Rahman (School of Data Science, Mathematics, and Informatics, IPB University)
Hwan-Seung Yong (Department of Computer Science and Engineering, Ewha Womans University, Republic of Korea)



Article Info

Publish Date
30 Jun 2026

Abstract

Sentiment analysis is widely used to extract user opinions from large-scale textual data. However, in practical settings, sentiment labels are often unavailable, making it difficult not only to perform sentiment classification but also to evaluate whether the predicted labels are reliable. This study proposes a two-stage framework for unsupervised sentiment analysis with model selection and semantic similarity evaluation. The dataset consists of Gemini app reviews, in which a labeled subset was used in the first stage to investigate the behavioral characteristics and predictive patterns of three sentiment analysis approaches: lexicon-based, transformer-based, and large language model (LLM)-based methods. The model with the most suitable performance was then selected and applied to predict sentiment labels for the remaining unlabeled data in the second stage. The predicted labels were further evaluated using embedding-based cosine similarity to assess semantic consistency within sentiment classes and separability between classes. The results show that the LLM-based method using Gemini 2.0 Flash achieved the best performance, with accuracy, balanced accuracy, and F1-score values above 0.91, followed by the transformer-based IndoBERT model, while the InSet lexicon-based method showed the weakest performance. In the second stage, semantic similarity evaluation revealed a high average intra-class similarity of 0.6699 and a low inter-class similarity of 0.1917, resulting in a similarity gap of 0.4783. These findings indicate that the proposed framework can support reliable sentiment prediction in largely unlabeled datasets by combining sample-based model selection with semantic validation.

Copyrights © 2026






Journal Info

Abbrev

ijsa

Publisher

Subject

Computer Science & IT Mathematics Other

Description

Indonesian Journal of Statistics and Its Applications (eISSN:2599-0802) (formerly named Forum Statistika dan Komputasi), established since 2017, publishes scientific papers in the area of statistical science and the applications. The published papers should be research papers with, but not limited ...