Classical statistical analysis of data can be complemented or replaced with data analysis based on machine learning. However, in certain disciplines, such as education research, studies are frequently limited to small datasets, which raises several questions regarding biases and coincidentally positive results. In this study, we present a refined approach for evaluating the performance of a binary classification based on machine learning for small datasets. The approach includes a non-parametric permutation test as a method to quantify the probability of the results generalising to new data. Furthermore, we found that a repeated nested cross-validation is almost free of biases and yields reliable results that are only slightly dependent on chance. Considering the advantages of several evaluation metrics, we suggest a combination of more than one metric to train and evaluate machine learning classifiers. In the specific case that both classes are equally important, the Matthews correlation coefficient exhibits the lowest bias and chance for coincidentally good results. The results indicate that it is essential to avoid several biases when analysing small datasets using machine learning.

Download full-text PDF

Source
http://www.ncbi.nlm.nih.gov/pmc/articles/PMC11108166PMC
http://journals.plos.org/plosone/article?id=10.1371/journal.pone.0301276PLOS

Publication Analysis

Top Keywords

machine learning
20
small datasets
16
refined approach
8
approach evaluating
8
binary classification
8
based machine
8
machine
5
learning
5
small
4
evaluating small
4

Similar Publications

Detection of Hepatitis C Virus Infection from Patient Sera in Cell Culture Using Semi-Automated Image Analysis.

Viruses

November 2024

Department of Infectious Diseases, Molecular Virology, Section Virus-Host Interactions, Heidelberg University, 69120 Heidelberg, Germany.

The study of hepatitis C virus (HCV) replication in cell culture is mainly based on cloned viral isolates requiring adaptation for efficient replication in Huh7 hepatoma cells. The analysis of wild-type (WT) isolates was enabled by the expression of SEC14L2 and by inhibitors targeting deleterious host factors. Here, we aimed to optimize cell culture models to allow infection with HCV from patient sera.

View Article and Find Full Text PDF

In this study, we introduce a novel approach that integrates interpretability techniques from both traditional machine learning (ML) and deep neural networks (DNN) to quantify feature importance using global and local interpretation methods. Our method bridges the gap between interpretable ML models and powerful deep learning (DL) architectures, providing comprehensive insights into the key drivers behind model predictions, especially in detecting outliers within medical data. We applied this method to analyze COVID-19 pandemic data from 2020, yielding intriguing insights.

View Article and Find Full Text PDF

Application of Machine Learning to Predict CO Emissions in Light-Duty Vehicles.

Sensors (Basel)

December 2024

Department of Computer Science, School of Computing and Engineering, University of Huddersfield, Queensgate, Huddersfield HD1 3DH, UK.

Climate change caused by greenhouse gas (GHG) emissions is an escalating global issue, with the transportation sector being a significant contributor, accounting for approximately a quarter of all energy-related GHG emissions. In the transportation sector, vehicle emissions testing is a key part of ensuring compliance with environmental regulations. The Vehicle Certification Agency (VCA) of the UK plays a pivotal role in certifying vehicles for compliance with emissions and safety standards.

View Article and Find Full Text PDF

Real-Time Freezing of Gait Prediction and Detection in Parkinson's Disease.

Sensors (Basel)

December 2024

School of Human Kinetics, Faculty of Health Sciences, University of Ottawa, Ottawa, ON K1N 6N5, Canada.

Freezing of gait (FOG) is a walking disturbance that can lead to postural instability, falling, and decreased mobility in people with Parkinson's disease. This research used machine learning to predict and detect FOG episodes from plantar-pressure data and compared the performance of decision tree ensemble classifiers when trained on three different datasets. Dataset 1 ( = 11) was collected in a previous study.

View Article and Find Full Text PDF

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!