IndoBERT-Based Sentiment Analysis of Indonesian Social Media Discourse on AI-Generated Images

Authors

  • Halvino Iqbal Nataprawira Universitas Pembangunan Jaya
  • Ida Nurhaida

DOI:

10.33395/sinkron.v10i3.16242

Keywords:

Artificial Intelligence, IndoBERT, Machine Learning, Sentiment Analysis, Social Media

Abstract

The rapid emergence of generative artificial intelligence has disrupted creative ecosystems, prompting widespread discourse across Indonesian social media. However, the exact sentiment structure of this public reaction remains empirically unmapped due to the contextual complexities of informal language. The objective of this research is to evaluate the efficacy of contextual language models by fine-tuning IndoBERT and benchmarking it against classical machine learning classifiers—including Complement Naive Bayes, Logistic Regression, and Support Vector Machine—for classifying social media sentiment. A multi-platform dataset comprising 2,981 Indonesian-language posts from X, Reddit, and YouTube was collected and manually annotated into positive, neutral, and negative classes. To address inherent class imbalance, Synthetic Minority Oversampling Technique was applied to classical models, while class-weighted loss and Masked Language Modeling augmentation were utilized for IndoBERT. Performance was evaluated using macro-averaged F1-score across five repeated stratified random splits. IndoBERT achieved a mean macro-F1 of 0.7131 ± 0.0180, outperforming the best classical baseline by approximately 0.12, demonstrating a pronounced advantage in resolving ambiguous neutral discourse. Negative sentiment heavily dominated the corpus at 61.8%, reflecting a prevailing critical stance toward AI-generated imagery concerning ethical and copyright issues. Furthermore, evaluation variance across random seeds exceeded variance from augmentation strategies, indicating test set composition is a major performance determinant. In conclusion, this study establishes a robust empirical baseline for Indonesian sentiment analysis, proving transformer architectures superior for nuanced public opinion mining.

GS Cited Analysis

Downloads

Download data is not yet available.

References

Alfat, L., Nasucha, M., Uddin, N., Salleh, K. A., & Baharun, N. (2023). Sentiment Classification of Indonesian Emotion Related to Vaccination Event using LSTM. 2023 IEEE World AI IoT Congress, AIIoT 2023, 825–829. https://doi.org/10.1109/AIIOT58121.2023.10174581

Asokere, M., Wusu, A., & Olabanjo, O. (2025). Twitter (X) as an electoral barometer: systematic evidence from sentiment analysis of Twitter data. International Journal of Information Technology 2025, 1–24. https://doi.org/10.1007/s41870-025-03039-1

Bello, A., Ng, S. C., & Leung, M. F. (2023). A BERT Framework to Sentiment Analysis of Tweets. Sensors 2023, Vol. 23, Page 506, 23(1), 506. https://doi.org/10.3390/s23010506

Bojanowski, P., Grave, E., Joulin, A., & Mikolov, T. (2017). Enriching Word Vectors with Subword Information. Transactions of the Association for Computational Linguistics, 5, 135–146. https://doi.org/10.1162/tacl_a_00051

Brewer, P. R., Cuddy, L., Dawson, W., & Stise, R. (2024). Artists or art thieves? media use, media messages, and public opinion about artificial intelligence image generators. AI & SOCIETY 2024 40:1, 40(1), 77–87. https://doi.org/10.1007/s00146-023-01854-3

Ceh-Varela, E., Shakya, S. R., & Imhmed, E. (2025). Challenges in Detecting Nuanced Sentiment with Advanced Models. Cloud Computing and Data Science, 6, 115–135. https://doi.org/10.37256/ccds.6220256316

Chawla, N. V., Bowyer, K. W., Hall, L. O., & Kegelmeyer, W. P. (2002). SMOTE: Synthetic Minority Over-sampling Technique. Journal of Artificial Intelligence Research, 16, 321–357. https://doi.org/10.1613/jair.953

Devlin, J., Chang, M. W., Lee, K., & Toutanova, K. (2018). BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. NAACL HLT 2019 - 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies - Proceedings of the Conference, 1, 4171–4186. https://arxiv.org/pdf/1810.04805

Dodge, J., Ilharco, G., Schwartz, R., Farhadi, A., Hajishirzi, H., & Smith, N. (2020). Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping. http://arxiv.org/abs/2002.06305

Dwi Cahya, L., Luthfiarta, A., Imanuel, J., Krisna, T., Winarno, S., & Nugraha, A. (2023). Improving Multi-label Classification Performance on Imbalanced Datasets Through SMOTE Technique and Data Augmentation Using IndoBERT Model. Jurnal Nasional Teknologi Dan Sistem Informasi, 9(3), 290–298. https://doi.org/10.25077/teknosi.v9i3.2023.290-298

Fadilah Arfat, M., Nurkholis, A., Kurniawan, I., Pagari Alam No, J. Z., Ratu, L., Kedaton, K., & Bandari Lampung, K. (2022). Analisis Sentimen Masyarakat Indonesia Terkait Vaksin Covid-19 Pada Media Sosial Twitter Menggunakan Metode Support Vector Machine (Svm). Jurnal Informatika: Jurnal Pengembangan IT, 7(2), 96–103. https://doi.org/10.30591/jpit.v7i2.3549

Gede Ari Rama, B., Julia Mahadewi, K., Kunci, K., Buatan, K., Cipta, H., & Hukum, S. (2023). Urgensi Pengaturan Artificial Intelligence (AI) Dalam Bidang Hukum Hak Cipta Di Indonesia. JURNAL RECHTENS, 12(2), 209–224. https://doi.org/10.56013/RECHTENS.V12I2.2395

Hu, L., Li, C., Wang, W., Pang, B., & Shang, Y. (2022). Performance Evaluation of Text Augmentation Methods with BERT on Small-sized, Imbalanced Datasets. Proceedings - 2022 IEEE 4th International Conference on Cognitive Machine Intelligence, CogMI 2022, 125–133. https://doi.org/10.1109/CogMI56440.2022.00027

Kanbach, D. K., Heiduk, L., Blueher, G., Schreiter, M., & Lahmann, A. (2023). The GenAI is out of the bottle: generative artificial intelligence from a business model innovation perspective. Review of Managerial Science 2023 18:4, 18(4), 1189–1220. https://doi.org/10.1007/s11846-023-00696-z

Khaqqi, F. N., Alfat, L., & Nurhaida, I. (2025). Sentiment Analysis of US-China Tariffs using IndoBERT and Economic Impact on Indonesia. Journal of Applied Informatics and Computing, 9(6), 3000–3011. https://doi.org/10.30871/JAIC.V9I6.11544

Korsakpaisarn, O., Noraset, T., Lapamnuaypol, J., & Jin’No, K. (2026). Optimizing BERT for Sentiment Classification of Amazon Product Reviews: A Study on Class Imbalance and Misclassification Analysis. 1–6. https://doi.org/10.1109/isai-nlp66160.2025.11320736

Laba, N. (2024). Engine for the imagination? Visual generative media and the issue of representation. Media, Culture and Society, 46(8), 1599–1620. https://doi.org/10.1177/01634437241259950

Mariam, S., & Nurhaida, I. (2025). Analisis Sentimen berbasis Deep Learning Terhadap Kesetaraan Gender di Bidang STEM: Perspektif dan Implikasinya. Edumatic: Jurnal Pendidikan Informatika, 9(1), 69–78. https://doi.org/10.29408/edumatic.v9i1.29071

Miyazaki, K., Murayama, T., Uchiba, T., An, J., & Kwak, H. (2024). Public perception of generative AI on Twitter: an empirical study based on occupation and usage. EPJ Data Science 2024 13:1, 13(1), 2-. https://doi.org/10.1140/epjds/s13688-023-00445-y

Nashrullah, M. A., Puspitaningayu, P., Agustin Tjahyaningtijas, H. P., Fransisca, Y., & Fikril Alan, M. B. (2025). BERT-based Transfer Learning for Two-Class Sentiment Detection in Indonesian X apps Comments. 139–143. https://doi.org/10.1109/ICVEE66651.2025.11281404

Saputra, R., Pristyanto, Y., & Fajri, I. N. (2025). Generative AI Image Sentiment Analysis on Social Media X using TF-IDF and FastText. Journal of Applied Informatics and Computing, 9(5), 2509–2520. https://doi.org/10.30871/JAIC.V9I5.10627

sindu, pande, Permana, A. A. J., & Wijaya, I. N. S. W. (2024). IDENTIFIKASI DAN NORMALISASI TEKS SLANG DENGAN FASTTEXT PADA TWITTER DALAM BAHASA INDONESIA. Jurnal Pendidikan Teknologi Dan Kejuruan, 21(1), 33–44. https://doi.org/10.23887/jptkundiksha.v21i1.66381

Søgaard, A., Ebert, S., Bastings, J., & Filippova, K. (2021). We Need To Talk About Random Splits. EACL 2021 - 16th Conference of the European Chapter of the Association for Computational Linguistics, Proceedings of the Conference, 1823–1832. https://doi.org/10.18653/v1/2021.eacl-main.156

Syahputra, R., Yanris, G. J., & Irmayani, D. (2022). SVM and Naïve Bayes Algorithm Comparison for User Sentiment Analysis on Twitter. Sinkron : Jurnal Dan Penelitian Teknik Informatika, 6(2), 671–678. https://doi.org/10.33395/sinkron.v7i2.11430

Wilie, B., Vincentio, K., Winata, G. I., Cahyawijaya, S., Li, X., Lim, Z. Y., Soleman, S., Mahendra, R., Fung, P., Bahar, S., & Purwarianti, A. (2020). IndoNLU: Benchmark and Resources for Evaluating Indonesian Natural Language Understanding. Proceedings of the 1st Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics and the 10th International Joint Conference on Natural Language Processing, AACL-IJCNLP 2020, 843–857. https://doi.org/10.18653/v1/2020.aacl-main.85

Downloads


Crossmark Updates

How to Cite

Nataprawira, H. I., & Nurhaida, I. (2026). IndoBERT-Based Sentiment Analysis of Indonesian Social Media Discourse on AI-Generated Images. Sinkron : Jurnal Dan Penelitian Teknik Informatika, 10(3), 1464-1475. https://doi.org/10.33395/sinkron.v10i3.16242