Skip to content
Open access

CVI-Validated Indo-Transformer Framework for Intelligent Cooperative Supervision

Aug 2026 · Jurnal RESTI (Rekayasa Sistem dan Teknologi Informasi) · 0 citations

TL;DR

It is demonstrated that CVI-validated ensemble GenAI can construct consistent labels for low-resource administrative texts and that IndoBERT provides the strongest and most stable generalization for cooperative supervision classification.

Abstract

Cooperative supervision reports contain complex narrative structures and overlapping administrative terminology, complicating automatic classification into governance, risk profile, financial performance, and capital adequacy. Reliable automation is particularly important for accelerating the analysis of supervisory findings while addressing limited labeled data and imbalanced categories. This study aimed to develop and externally evaluate a text-classification framework combining quantitatively validated Generative Artificial Intelligence (GenAI) labeling with conventional and Transformer-based models. Data comprised 294 preprocessed sentences collected from the Department of Cooperatives, Small and Medium Enterprises, Industry, and Trade of Semarang Regency during 2023–2025. Few-shot annotations were generated using ChatGPT, Gemini, Perplexity, and DeepSeek, and three-model combinations were evaluated using the Content Validity Index (CVI); majority voting from the best combination established ground truth. TF-IDF with Logistic Regression and Support Vector Machine served as baselines, whereas IndoBERT and IndoRoBERTa represented contextual models. Performance was assessed through stratified five-fold cross-validation and external testing on 58 unseen sentences. ChatGPT–Gemini–Perplexity achieved the highest Scale-Level CVI of 0.898. IndoBERT obtained the best cross-validated F1-score of 0.9099, exceeding IndoRoBERTa (0.8217), Logistic Regression (0.8004), and SVM (0.7863). On unseen data, IndoBERT retained an F1-score of 0.862, compared with 0.759 for IndoRoBERTa. These findings demonstrate that CVI-validated ensemble GenAI can construct consistent labels for low-resource administrative texts and that IndoBERT provides the strongest and most stable generalization for cooperative supervision classification. The framework offers a practical basis for scalable annotation and reliable automated support for evidence-based supervisory decision-making.

Read PDF

Similar papers

Open access 2026

Integrating Transformer-based and Embedding Models into Rasa NLU for Vietnamese University Support System

This paper presents a Vietnamese university support chatbot developed using the Rasa Natural Language Understanding (NLU) framework, integrating Transformer-based and embedding models, including PhoBERT, FastText, Multilingual BERT (mBERT), and additional baseline methods such as Support Vector Machine (SVM) and Naive...

Le Ba Cuong, Le Anh Tien, Huong Van Pham · 0 citations
Review Open access Jul 2026

Cross-Domain Faithfulness Evaluation of SHAP and Attention-Based Explanations in Transformer NLP Models

The results demonstrate that superior predictive performance does not necessarily correspond to higher explanation faithfulness or stronger cross-domain stability, and highlight the importance of jointly evaluating predictive performance, explanation faithfulness, and explanation robustness when developing trustworthy...

Dony Bahtera Firmawan, B. Darnoto · 1 citation
Open access 2026

Advancing Machine-generated Text Detection: A Comprehensive Evaluation of Transformer-based Models

Test set results show that Decoding-Enhanced Bert with Disentangled Attention (DeBERTa) achieves the highest macro F1 − Score of 85.48%, surpassing the previously top-ranked Multi-Task Learning (MTL) system, which attains a macro F1 of 83.07%.

Batyr Sharimbayev, S. Kadyrov · 0 citations
Preprint Aug 2026

LabelFusion-TS: Fusing Large Language Models, Transformer Encoders, and Financial Time Series for Monetary-Policy Stance Classification

It is taken as initial evidence for market time series as an input modality in financial text classification on the task of classifying sentences from Federal Reserve communication as hawkish, dovish, or neutral.

Michael Schlee, Fabian Lukassen, Christoph Weisser · 0 citations
Review Open access Aug 2026

Evaluating GPT-4o-based Data Augmentation for Imbalanced Multiclass Sentiment Classification of GoPay Reviews Using IndoBERT-LoRA

This study evaluated GPT-4o-based data augmentation for imbalanced multiclass sentiment classification of GoPay user reviews using IndoBERT-LoRA. The main problem addressed in this study was the limited representation of minority sentiment classes, particularly the neutral class, which could reduce the model’s ability...

M. Yusran, F. Afendi, Anwar Fitrianto · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.