Skip to content
Data Science & information systems International Journal of Advances in Data and Information Systems
Open access E-ISSN 2721-3056 Acceptance rate: 28%

Machine Learning-Based Prediction of Divorce Verdicts Using Posita Data and Imbalanced Data Handling: A Case Study in Padang Sidempuan

Authors

  • Rina Rahmadini School of Interdisciplinary Management and Technology, Institut Teknologi Sepuluh Nopember Surabaya, Indonesia
  • Bagus Jati Santoso School of Interdisciplinary Management and Technology, Institut Teknologi Sepuluh Nopember Surabaya, Indonesia

DOI:

https://doi.org/10.59395/ijadis.v6i2.1405

Keywords:

Machine Learning, Court Verdict Prediction, Divorce Case Analysis, Posita Data, Imbalanced Data Handling

Abstract

This study aims to develop a predictive model for divorce verdicts ("Granted" or "Rejected") in the Religious Courts of Indonesia using machine learning techniques. The dataset consists of 2,026 finalized divorce cases from the Religious Court of Padang Sidempuan between 2018 and 2025, incorporating structured variables and posita—narrative texts describing the plaintiff’s reasons for divorce. Keyword-based feature extraction was applied to transform these texts into interpretable indicators. To handle class imbalance, Synthetic Minority Over-sampling Technique (SMOTE) was implemented on the training data. Six classical machine learning algorithms were evaluated: Decision Tree, Naïve Bayes, K-Nearest Neighbors, Random Forest, LightGBM, and XGBoost. Performance was measured using accuracy, precision, recall, F1-score, F2-score, and AUC. The results indicate that Naïve Bayes achieved the highest recall (100%) for the “Granted” class, while LightGBM and XGBoost demonstrated the most balanced performance across both classes. Feature importance analysis revealed that mediation outcomes, domestic violence, and economic hardship were among the most influential factors in determining verdicts. The study highlights the applicability of interpretable machine learning in legal decision support and discusses limitations such as the single-court scope and challenges in predicting minority class outcomes. Future work may explore multi-jurisdictional data, deep learning approaches, and domain-specific embeddings for enhanced performance.

1311 692

Downloads

Download data is not yet available.

References

[1] Badan Pusat Statistik, Statistik Indonesia 2023. Jakarta: BPS, 2023. [Online]. Available: https://www.bps.go.id

[2] Undang-Undang Republik Indonesia Nomor 3 Tahun 2006 tentang Perubahan atas Undang-Undang Nomor 7 Tahun 1989 tentang Peradilan Agama.

[3] S. Nahavandi, Industry5.0a human centric solution, Sustainability, vol.11, no.16, p.4371, 2019. doi: 10.3390/su11164371 DOI: https://doi.org/10.3390/su11164371

[4] M. Garca, J. Rodriguez, A. Sanchez, and L. Torres, "Big data and predictive analytics: A systematic review of applications," Artificial Intelligence Review, vol. 57, no. 4, pp. 201252, 2024, doi: 10.1007/s10462-024-10811-5 DOI: https://doi.org/10.1007/s10462-024-10811-5

[5] H. A. Almuzaini and A. M. Azmi, "TaSbeeb: A judicial decision support system based on deep learning framework," Journal of King Saud University Computer and Information Sciences, vol. 35, no. 8, p. 101695, 2023, doi: 10.1016/j.jksuci.2023.10169 DOI: https://doi.org/10.1016/j.jksuci.2023.101695

[6] A. Sohail, H. Khalid, F. Khalid, dan M. Akram, Implementation of machine learning algorithm on factors affecting divorce rate, International Journal of Scientific & Engineering Research, vol. 9, no. 5, pp. 602607, 2018 DOI: https://doi.org/10.1109/ICEET1.2018.8338618

[7] M. Medvedeva, M. Vols, dan M. Wieling, Using machine learning to predict decisions of the European Court of Human Rights, Artificial Intelligence and Law, vol. 28, no. 2, pp. 237266, 2020, doi: 10.1007/s10506-019-09255-y DOI: https://doi.org/10.1007/s10506-019-09255-y

[8] N. Aimran, R. A. Othman, and A. S. Ahmad, "Prediction of Malaysian women divorce using machine learning techniques," Malaysian Journal of Computing, vol. 7, no. 2, pp. 491502, 2022, doi: 10.24191/mjoc.v7i2.17077. DOI: https://doi.org/10.24191/mjoc.v7i2.17077

[9] A. Bramantoro dan I. Virdyna, Classification of divorce causes during the COVID-19 pandemic using convolutional neural networks, PeerJ Computer Science, vol. 8, p. e998, 2022, doi: 10.7717/peerj-cs.998 DOI: https://doi.org/10.7717/peerj-cs.998

[10] A. B. Dina, R. Sarno, R. N. E. Anggraini, A. T. Haryono, and A. F. Septiyanto, "Comparison of oversampling techniques in prediction judicial decisions of divorce trials in family courts," in Proc. 2024 Int. Conf. on Information Technology Research and Innovation (ICITRI), Surabaya, Indonesia, 2024, pp. 1318, doi: 10.1109/ICITRI62858.2024.10699016. DOI: https://doi.org/10.1109/ICITRI62858.2024.10699016

[11] A. S. Maarif and D. H. Firmansyah, "Fulfilling childrens rights through post-divorce court decisions: A case study of religious court verdicts in Indonesia," Ahwal: Jurnal Hukum Keluarga Islam, vol. 16, no. 1, pp. 123140, 2023, doi: 10.14421/ahwal.2023.16108 DOI: https://doi.org/10.14421/ahwal.2023.16108

[12] D. Kusnandar and F. Rahma, Optimizing Legal Protection for Divorce Outside of Court: Study of the Need for Divorce Isbat in the Indonesian Legal System, Indonesian Journal of Islamic Law, vol. 6, no. 2, pp. 7388, Jun. 2023, doi: 10.35719/ijil.v6i1. DOI: https://doi.org/10.35719/ijil.v6i2.2010

[13] M. Mujahid, Y. T. Shah, and M. Khan, Machine learning with class imbalance: A review, Journal of Big Data, vol. 11, no. 1, art. no. 71, 2024, doi: 10.1186/s40537-024-00943-4 DOI: https://doi.org/10.1186/s40537-023-00851-z

[14] T. Subrata, Title and judiciary power based on Law Number 48 of 2009, *Legal Brief*, vol. 11, no. 5, pp. 28752881, 2022, doi: 10.35335/legal. DOI: https://doi.org/10.35335/legal

[15] Direktorat Jenderal Badan Peradilan Agama, "Sistem Informasi Penelusuran Perkara (SIPP)," Mahkamah Agung Republik Indonesia, [Online]. Available: https://sipp.mahkamahagung.go.id

[16] S. Matharaarachchi, M. Domaratzki, and S. Muthukumarana, Enhancing SMOTE for imbalanced data with abnormal minority instances, Machine Learning with Applications, vol. 19, art. no. 100597, 2024, doi: 10.1016/j.mlwa.2024.100597 DOI: https://doi.org/10.1016/j.mlwa.2024.100597

[17] M. OwusuAdjei, J. Ben HayfronAcquah, T. Frimpong, and G. AbdulSalaam, "Imbalanced class distribution and performance evaluation metrics: A systematic review of prediction accuracy for determining model performance in healthcare systems," PLOS Digital Health, vol. 2, no. 11, p. e0000290, Nov. 2023, doi: 10.1371/journal.pdig.0000290 DOI: https://doi.org/10.1371/journal.pdig.0000290

[18] M. Saarela and S. Jauhiainen, Comparison of feature importance measures as explanations for classification models, SN Applied Sciences, vol. 3, no. 6, p. 414, 2021, doi: 10.1007/s42452-021-04148-9 DOI: https://doi.org/10.1007/s42452-021-04148-9

[19] Y. Zhang, C. Li, Y. Sheng, J. Ge, and B. Luo, Judicial intelligent assistant system: Extracting events from Chinese divorce cases to detect disputes for the judge, Expert Systems, vol. 41, no. 7, e13540, Jan. 2024, doi: 10.1111/exsy.13540 DOI: https://doi.org/10.1111/exsy.13540

[20] M. Alshamrani, M. Almaiah, and S. Al-Khalifa, A machine learning-based approach to predict user satisfaction in mobile government services, PeerJ Computer Science, vol. 9, e2131, 2023, doi: 10.7717/peerj-cs.2131 DOI: https://doi.org/10.7717/peerj-cs.2131

[21] M. Altalhan, A. Algarni, and M. Turki-Hadj Alouane, Imbalanced data problem in machine learning: A review, IEEE Access, vol. 13, pp. 1339813424, Jan. 2025, doi: 10.1109/ACCESS.2025.3531662 DOI: https://doi.org/10.1109/ACCESS.2025.3531662

[22] I. Domormienye and N. Jere, A survey of decision trees: Concepts, algorithms, and applications, IEEE Access, vol. 12, pp. 86716-86727, 2024, doi: 10.1109/ACCESS.2024.3416838. DOI: https://doi.org/10.1109/ACCESS.2024.3416838

[23] S. S. Prasetiyowati and Y. Sibaroni, Unlocking the potential of Naive Bayes for spatio temporal classification: a novel approach to feature expansion, Journal of Big Data, vol. 11, no. 106, 2024, doi: 10.1186/s40537-024-00958-x DOI: https://doi.org/10.1186/s40537-024-00958-x

[24] R. K. Halder, M. N. Uddin, M. A. Uddin, S. Aryal, and A. Khraisat, Enhancing K-nearest neighbor algorithm: a comprehensive review and performance analysis of modifications, Journal of Big Data, vol. 11, no. 113, 2024. doi: 10.1186/s40537-024-00973-y DOI: https://doi.org/10.1186/s40537-024-00973-y

[25] L. Barreada, P. Dhiman, D. Timmerman, A.-L. Boulesteix, and B. Van Calster, Understanding overfitting in random forest for probability estimation: a visualization and simulation study, Diagnostic and Prognostic Research, vol. 8, no. 14, 2024, doi: 10.1186/s41512-024-00177-1 DOI: https://doi.org/10.1186/s41512-024-00177-1

[26] T. O. Ometehinwa, D. O. Oyewola, and E. G. Moung, Optimizing the light gradient-boosting machine algorithm for an efficient early detection of coronary heart disease, Informatics and Health, vol. 1, p.70-81, 2024, doi:10.1016/j.infoh.2024.06.001 DOI: https://doi.org/10.1016/j.infoh.2024.06.001

[27] J. Jinbo, L. Yufu, dan M. Haitao, Handling missing data of using the XGBoost based multiple imputation by chained equations regression method, Frontiers in Artificial Intelligence., vol. 8, p.1553220, Apr. 2025, doi:10.3389/frai.2025.1553220 DOI: https://doi.org/10.3389/frai.2025.1553220

[28] X. He, K. Zhao, and X. Chu, AutoML: A survey of the state-of-the-art, KnowledgeBased Systems, vol. 212, p. 106622, 2021, doi:10.1016/j.knosys.2020.106622 DOI: https://doi.org/10.1016/j.knosys.2020.106622

[29] G. M. Foody, Challenges in the real world use of classification accuracy metrics: From recall and precision to the Matthews correlation coefficient, PLOS ONE, 2023, doi: 10.1371/journal.pone.0291908 DOI: https://doi.org/10.1371/journal.pone.0291908

[30] J. Li, Area under the ROC Curve has the most consistent evaluation for binary classification, PLOS ONE, 2023, doi: 10.1371/journal.pone.0316019 DOI: https://doi.org/10.1371/journal.pone.0316019

[31] A. Kumar, and J. W. Taylor, Feature importance in the age of explainable AI: Case study of detecting fake news & misinformation via a multi-modal framework, European Journal of Operational Research, vol.317, pp.401413, 2023, doi: 10.1016/j.ejor.2023.10.003 DOI: https://doi.org/10.1016/j.ejor.2023.10.003

Downloads

Published

2025-08-30

How to Cite

[1]
R. Rahmadini and B. J. Santoso, “Machine Learning-Based Prediction of Divorce Verdicts Using Posita Data and Imbalanced Data Handling: A Case Study in Padang Sidempuan”, International Journal of Advances in Data and Information Systems, vol. 6, no. 2, pp. 460–478, Aug. 2025, doi: 10.59395/ijadis.v6i2.1405.

Share



Plum Analytics


Similar Articles

1-10 of 182

You may also start an advanced similarity search for this article.