Skip to content
Open access

Dominant-learner adaptive mixing for concrete compressive strength prediction

Aug 2026 · Frontiers in Materials · 0 citations · 60 references

Abstract

Accurate prediction of concrete compressive strength is essential for mixture design, quality control, and the broader use of supplementary cementitious materials in low-carbon construction. Fly ash concrete is particularly challenging to model because its strength development is affected by nonlinear interactions among binder composition, water–binder relationships, admixture dosage, and material characteristics. To address this problem, this study proposes a Dominant Learner with Adaptive Mixing (DLAM) framework for data-driven strength prediction. DLAM uses inner cross-validation to identify the most reliable learner from a pool of machine learning models and introduces a validation-controlled Ridge calibration step to exploit complementary information among candidate predictions. The calibration branch is adopted only when it improves the inner-validation root mean squared error (RMSE), thereby reducing the risk of unnecessary model combination and performance degradation. The framework is evaluated using a leakage-free repeated outer/inner validation protocol on a fly ash concrete dataset and is further examined on an independent public concrete strength dataset. DLAM is compared with individual learners, adaptive model-averaging baselines, and Stacking. The results show that DLAM achieves the lowest mean RMSE among the focused comparators on both datasets, with a clear improvement on the external dataset and a more modest gain on the fly ash dataset. These findings demonstrate that validation-controlled calibration provides a transparent and robust way to enhance machine-learning-based concrete strength prediction, especially when different learners capture complementary aspects of the mixture–strength relationship.

Read PDF

Similar papers

Open access 2026

Machine Learning Approaches for Predicting Compressive Strength of Concrete: A Comparative Performance Analysis

Accurate prediction of concrete compressive strength is essential for effective mix design, quality control, and structural performance assessment. Conventional empirical models often exhibit limited accuracy due to the complex and nonlinear interactions among concrete constituents. This study investigates the applicability of several machine learning models for predicting the compressive strength of concrete using a publicly available experimental dataset comprising 1030 concrete mixtures. Linear regression was adopted as a baseline model and compared with support vector regression, random forest regression, and artificial neural networks. The performance of machine learning models was meticulously assessed using the coefficient of determination, root mean square error, and mean absolute error. Additionally, the models underwent five-fold cross-validation to evaluate their robustness and generalization capabilities. The results unambiguously demonstrate that machine learning models significantly outperform linear regression models. Cross-validation results confirm the stability and reliability of the developed models. Feature importance analysis reveals that curing age and cement content are the most influential parameters affecting compressive strength, followed by water content, which is consistent with established concrete material behavior. The findings demonstrate that machine learning models, particularly random forest regression, can serve as effective supporting tools for preliminary concrete mix design and performance evaluation.

S. Rouabah · 0 citations
Open access Jul 2026

Data-driven modeling of compressive strength in sustainable self-compacting concrete incorporating recycled aggregates using ensemble learning techniques

Abstract This study develops a robust framework for estimating the compressive strength of self-compacting concrete (SCC) incorporating recycled aggregates using supervised machine learning (ML) techniques. A comprehensive experimental database comprising 582 concrete mix designs was used, encompassing diverse input variables including binder content, water, coarse and fine aggregates, recycled aggregate proportion, superplasticizer dosage, and curing time. Seven ML algorithms—XGBoost, CatBoost, AdaBoost, Extra Trees, Bagging Regressor, K-Nearest Neighbors, and Radius Neighbors—were systematically trained using a stratified 70/15/15 data split and optimized via grid search with five-fold cross-validation. Model performance was evaluated using coefficient of determination (R 2), root mean squared error, and MAE across training, validation, and testing datasets. Among all models, XGBoost demonstrated the highest accuracy, achieving an average R 2 of 0.9799, RMSE of 2.87 MPa, and mean absolute error of 1.97 MPa. The Permutation Feature Importance analysis revealed that binder content, water, and coarse aggregate were the most influential predictors of strength. This study confirms that ensemble ML models, particularly XGBoost, can reliably predict the compressive strength of SCC with recycled aggregates, while offering transparent insights into material behavior. The results provide a valuable tool for sustainable mix design optimization and practical implementation in eco-efficient concrete construction.

A. Khan, M. D. Rasheed, Muhammad Huzaifa Naveed et al. · 0 citations
Open access Jul 2026

Machine Learning Prediction of Concrete Compressive Strength: Model Comparison, CatBoost Optimization, and SHAP Interpretation

A comparative framework evaluating nine regression algorithms using the UCI Concrete Compressive Strength dataset, jointly integrating correlation-corrected statistical validation, multi-model Bayesian optimization, and domain-informed feature engineering with SHAP interpretation, rarely combined in prior concrete-strength studies.

Musthafa 'Abduh Fakhruddin, Sri Winarno, Acun Kardianawati · 0 citations
Open access Aug 2026

Explainable Random Forest Framework for Predicting Compressive Strength of Sustainable Concrete Incorporating Industrial Waste Materials

Compressive strength is the single most important design parameter governing the safety, serviceability, and economy of concrete structures, yet its determination through standard 7-, 14-, or 28-day destructive cylinder/cube testing is slow, costly, and unable to assess concrete already cast in place. This study develops and evaluates a Random Forest (RF) regression model to predict the compressive strength of concrete directly from eight standard mix-design parameters — cement, blast furnace slag, fly ash, water, superplasticizer, coarse aggregate, fine aggregate, and curing age — using Yeh's (1998) benchmark dataset of 1,030 experimentally tested concrete mixtures. Following data cleaning, exploratory correlation analysis, an 80:20 train-test split, and five-fold GridSearchCV hyperparameter tuning, the optimized Random Forest model is benchmarked against Linear Regression, Ridge Regression, and Support Vector Regression using the coefficient of determination (R²), Root Mean Squared Error (RMSE), and Mean Absolute Error (MAE). The Random Forest model achieves the strongest predictive performance of the models tested, substantially outperforming the linear baselines and confirming that concrete strength development is governed by non-linear interactions among mix constituents. Feature importance analysis further shows that curing age and cement content are the dominant predictors, while water content exerts a clear negative influence consistent with Abrams' Law, and coarse/fine aggregates contribute comparatively little, consistent with their role as largely inert fillers. These findings demonstrate that Random Forest regression offers a fast, accurate, and interpretable, non-destructive alternative to conventional strength testing, with practical value for mix-design optimization, quality control, and early-stage structural decision-making.

M. Selvakumar, S. Geetha, P. K. Kumar et al. · 0 citations
Open access Jul 2026

Integrated Prediction Model for Normal and Recycled Aggregate Concrete Strength Using Ensemble Learning Techniques

Recycled aggregate concrete (RAC) is a sustainable alternative construction material to reduce natural resource exploitation and manage construction and demolition waste. However, predicting the mechanical performance of RAC remains a challenge due to the high variability of recycled aggregate properties. The purpose of this study is to develop a machine learning model to predict the compressive strength of recycled aggregate-based concrete and compare its performance with normal concrete. The dataset used consists of 2165 samples (1600 normal concrete and 565 recycled aggregate concrete) collected from various scientific publications. Three tree-based machine learning algorithms (Random Forest, XGBoost, and LightGBM) were implemented and optimized using RandomizedSearchCV with 5-fold cross-validation. The results showed that LightGBM provided the best performance with R² = 0.92, MAE = 2.45 MPa, and RMSE = 3.52 MPa on the test set. This model is able to predict the compressive strength of normal concrete (R² = 0.92) and recycled aggregate concrete (R² = 0.91) with almost the same accuracy, indicating strong generalization. Feature importance analysis revealed that curing age, cement content, and water content are the most important factors in compressive strength prediction, while for RAC, recycled aggregate water absorption (WRCA) also makes a significant contribution. Error analysis shows that residuals are random and normally distributed without systematic bias. This model can reliably predict concrete compressive strength in the range of 20-60 MPa with an average error of ±3-4 MPa and can be integrated into mix proportioning design software to improve the efficiency of the design process and support the use of sustainable construction materials.

Suji’at, Eko Wahyu Abryandoko, Ocha Silvia Kencana et al. · 0 citations
Conference Open access Jul 2026

Machine Learning-Based Prediction of Compressive Strength in Basalt Fiber Reinforced Concrete

Accurate prediction of the mechanical strength of Basalt Fiber Reinforced Concrete (BFRC) is critical for structural design, safety assessment, and the advancement of sustainable infrastructure in civil engineering. Traditional prediction methods often fail to capture the nonlinear relationships between BFRC mix proportions and resulting strength characteristics, leading to unreliable estimations. To address this limitation, this study proposes the Optimized Moment Balanced Machine (OMBM), an advanced machine learning model developed to improve the predictive accuracy of BFRC strength parameters. The model was trained and evaluated using key input features, including cement content, silica fume, fly ash, superplasticizer, water, aggregate composition, and fiber property parameters. The performance of the OMBM was benchmarked against four established machine learning models, such as Least Squares Support Vector Machine (LSSVM), Backpropagation Neural Network (BPNN), K-Nearest Neighbors (KNN), and Linear Regression (LR). Results from ten-fold cross-validation show that OMBM consistently outperforms the comparison models across five evaluation metrics. It achieved the lowest RMSE (2.411), MAE (1.788), and MAPE (4.08%), along with the highest values for correlation coefficient (R = 0.978), and coefficient of determination (R2 = 0.956). Furthermore, the OMBM achieved a Reference Index (RI) score of 1.000, which confirms its position as the leading predictive model within this comparative framework. These results confirm the robustness and reliability of the proposed OMBM model, making it a highly effective tool for accurate strength prediction of BFRC. This approach offers significant potential for the advancement of sustainable infrastructure by enabling more accurate and efficient use of concrete materials.

R. R. Khasani, Ferry Hermawan, Yuliana Usman · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.