INVESTIGASI HATE SPEECH PADA PLATFORM "X" MENGGUNAKAN METODE HYBRID INDOBERT–GRAPH ATTENTION NETWORK

Authors

  • Weslie Austin University of Bunda Mulia
  • Puguh Hiskiawan University of Bunda Mulia image/svg+xml

DOI:

https://doi.org/10.33480/inti.v21i1.8699

Keywords:

Ablation Study, Graph Attention Network, Hate Speech, Hybrid IndoBERT-GAT

Abstract

The rapid growth of social media, particularly the X (formerly Twitter) platform, has made it a primary space for public discourse in Indonesia, yet its openness has also become a systematic loophole for the spread of hate speech that threatens social cohesion. Conventional text-based detection models fail to capture hidden linguistic nuances such as slang and regional euphemisms, and they disregard the social dimension of hate propagation reflected in user interaction patterns. This study proposes a hybrid IndoBERT-GAT architecture that integrates textual semantic representation with social graph structure into a single unified framework. The research method employed 16,646 government-related tweets collected via Apify crawling, labeled through a hybrid approach (manual annotation and pseudo-labeling), then represented through IndoBERT embeddings for textual features and a Graph Attention Network to model a heterogeneous tweet-user graph, before being combined via a feature fusion mechanism and evaluated through an ablation study across four model scenarios. The results show that the full Hybrid IndoBERT-GAT model achieved the highest test F1-score (65%), outperforming the same architecture without metadata (61.8%), the GAT+metadata-only baseline (42.2%), and the MLP+metadata baseline (39.9%), demonstrating that IndoBERT's semantic representation is the dominant contributor while graph structure and metadata serve as complementary signals, although the model still tends to overpredict hate speech in politically sarcastic content that uses sharp language without genuinely hateful intent.

Downloads

Download data is not yet available.

References

Ali, R., Farooq, U., Arshad, U., Shahzad, W., & Beg, M. O. (2022). Hate speech detection on Twitter using transfer learning. Computer Speech & Language, 74, 101365. https://doi.org/10.1016/j.csl.2022.101365

Aliansi Jurnalis Independen (AJI) Indonesia; Monash University Indonesia. (2024). Report on Hate Speech Monitoring in the 2024 Indonesia Regional Elections. https://www.monash.edu/indonesia/news/report-on-hate-speech-monitoring-in-the-2024-indonesia-regional-elections

Benomar, A., Zarour, E., Létourneau-Guillon, L., & Raymond, J. (2023). Measuring Interrater Reliability. Radiology, 309(3). https://doi.org/10.1148/radiol.230492

Cinelli, M., Cresci, S., Quattrociocchi, W., Tesconi, M., & Zola, P. (2022). Coordinated inauthentic behavior and information spreading on Twitter. Decision Support Systems, 160, 113819. https://doi.org/10.1016/j.dss.2022.113819

Hakim, A. N., Sibaroni, Y., & Prasetyowati, S. S. (2024). Detection of Hate-Speech Text on Indonesian Twitter Social Media Using IndoBERTweet-BiLSTM-CNN. 2024 12th International Conference on Information and Communication Technology (ICoICT), 374–381. https://doi.org/10.1109/ICoICT61617.2024.10698615

Hiskiawan, P., Geasela, Y. M., Heryanto, H., Stephanie, E., Ardianti, M., & Sukarno, F. A. (2025). Trustworthy Data Science Framework for Non-Invasive Nutritional Screening Using Computer Vision. 2025 International Conference on Informatics, Multimedia, Cyber and Information System (ICIMCIS), 1743–1748. https://doi.org/10.1109/ICIMCIS68501.2025.11327268

Hiskiawan, P., Indarwan, M. S., Ho, C. A., & Sutojo, K. S. B. (2026). Pemberdayaan Siswa SMK DKV melalui Praktikum Interaktif Natural Language Processing. PengabdianMu: Jurnal Ilmiah Pengabdian kepada Masyarakat, 11(3), 949–956. https://doi.org/10.33084/pengabdianmu.v11i3.11700

Hiskiawan, P., Sari, M. K., Siregar, R. E., Valencia, N. A., & Wijayanti, T. P. (2026). A Deep Learning Data Fusion Approach for Prostate MRI Zonal Segmentation Using Dual-Channel U-Net with T2 and ADC Images. 2026 International Conference on Current Research in Artificial Intelligence and Data Science (ICCRAIDS), 1–7. https://doi.org/10.1109/ICCRAIDS67816.2026.11519671

Ibrohim, M. O., & Budi, I. (2023). Hate speech and abusive language detection in Indonesian social media: Progress and challenges. Heliyon, 9(8), e18647. https://doi.org/10.1016/j.heliyon.2023.e18647

Jahan, M. S., & Oussalah, M. (2023). A systematic review of hate speech automatic detection using natural language processing. Neurocomputing, 546, 126232. https://doi.org/10.1016/j.neucom.2023.126232

Kusuma, J. F., & Chowanda, A. (2023). Indonesian Hate Speech Detection Using IndoBERTweet and BiLSTM on Twitter. JOIV : International Journal on Informatics Visualization, 7(3), 773–780. https://doi.org/10.30630/joiv.7.3.1035

Miao, Z., Chen, X., Wang, H., Tang, R., Yang, Z., Huang, T., & Tang, W. (2024). Detecting Offensive Language Based on Graph Attention Networks and Fusion Features. IEEE Transactions on Computational Social Systems, 11(1), 1493–1505. https://doi.org/10.1109/TCSS.2023.3250502

Min, B., Ross, H., Sulem, E., Veyseh, A. P. Ben, Nguyen, T. H., Sainz, O., Agirre, E., Heintz, I., & Roth, D. (2024). Recent Advances in Natural Language Processing via Large Pre-trained Language Models: A Survey. ACM Computing Surveys, 56(2), 1–40. https://doi.org/10.1145/3605943

Mufva, P. A., Chandra, K. H., Aji, K. F., Iswanto, I. A., & Joddy, S. (2025). Performance comparison of deep learning approaches for Indonesian twitter hate speech detection using IndoBERTweet embedding. Procedia Computer Science, 269, 1663–1671. https://doi.org/10.1016/j.procs.2025.09.109

Nagar, S., Gupta, S., Bahushruth, C. S., Barbhuiya, F. A., & Dey, K. (2022). Hate Speech Detection on Social Media Using Graph Convolutional Networks (hlm. 3–14). https://doi.org/10.1007/978-3-030-93413-2_1

Pamungkas, E. W., & Chiril, P. (2025). Ngalawan Ujaran Sengit: hate speech detection in indonesian code-mixed social media data. Language Resources and Evaluation, 59(3), 2387–2414. https://doi.org/10.1007/s10579-025-09810-x

Perwira Joan Dwitama, A., Hatta Fudholi, D., & Hidayat, S. (2023). Indonesian Hate Speech Detection Using Bidirectional Long Short-Term Memory (Bi-LSTM). Jurnal RESTI (Rekayasa Sistem dan Teknologi Informasi), 7(2), 302–309. https://doi.org/10.29207/resti.v7i2.4642

Rawat, A., Kumar, S., & Samant, S. S. (2024). Hate speech detection in social media: Techniques, recent trends, and future challenges. WIREs Computational Statistics, 16(2). https://doi.org/10.1002/wics.1648

Vrahatis, A. G., Lazaros, K., & Kotsiantis, S. (2024). Graph Attention Networks: A Comprehensive Review of Methods and Applications. Future Internet, 16(9), 318. https://doi.org/10.3390/fi16090318

Winson, W., & Hiskiawan, P. (2026). An Applied Data Science Approach for Detecting Depression Symptoms in Indonesian Social Media Text Using Transformer Models. Journal of Applied Informatics and Computing, 10(3), 2115–2127. https://doi.org/10.30871/jaic.v10i3.12644

Downloads

Published

2026-09-03

How to Cite

INVESTIGASI HATE SPEECH PADA PLATFORM "X" MENGGUNAKAN METODE HYBRID INDOBERT–GRAPH ATTENTION NETWORK. (2026). INTI Nusa Mandiri, 21(1), 193-200. https://doi.org/10.33480/inti.v21i1.8699