DOI: 10.1021/acsomega.6c05807 ISSN: 2470-1343

A Leakage-Aware Benchmark Study of Machine Learning Models for Deep Eutectic Solvent Property Prediction

Hakim Faraji, Julio Brito Santana, Ruth Rodríguez-Ramos, Bárbara Socas-Rodríguez, Antonio V. Herrera Herrera

Abstract

Predicting the physicochemical properties of deep eutectic solvents (DESs) remains challenging due to the large combinatorial design space and the complex, composition- and temperature-dependent interactions governing their behavior. While machine learning (ML) has been widely applied to DES property prediction, reported performance is often sensitive to data set structure, feature representation, and validation design, raising questions about the reliability and transferability of existing models. In this work, we present a systematic evaluation of descriptor-based ML models for DES property prediction using a curated and high-confidence data set spanning 2003–2026. A unified feature representation combining molecular descriptors, molar composition, and temperature is employed, and model performance is assessed under a hierarchy of validation protocols designed to control for data leakage and progressively increase extrapolation difficulty. The results show that predictive performance is strongly dependent on the validation design. Density and refractive index exhibit relatively stable behavior, while surface tension shows moderate predictability. In contrast, electrical conductivity and viscosity display a strong dependence on temperature and limited contribution from descriptor-based features. Under extrapolative validation, performance for these properties deteriorates substantially, indicating limited transferability. Additional analyses based on feature-space distance and similarity-based baselines indicate that prediction errors are strongly associated with training-domain proximity and that local similarity in the current feature space is insufficient for reliable prediction outside observed data regions. These findings suggest that the primary limitation arises from the representational capacity of static descriptor-based features rather than model choice alone. Overall, this study provides a structured and reproducible assessment of the conditions under which descriptor-based ML models can be expected to succeed or fail in DES systems and highlights the importance of rigorous, leakage-aware evaluation in data-driven chemical modeling.

More from our Archive