ORIGINAL RESEARCH article

Front. Digit. Health, 01 October 2025

Sec. Connected Health

Volume 7 - 2025 | https://doi.org/10.3389/fdgth.2025.1612915

Heart disease prediction with a feature-sensitized interpretable framework for the Internet of Medical Things sensors

  • 1. Balaji Institute of Modern Management, Sri Balaji University, Pune, India

  • 2. Department of Computer Science, Loyola College, Chennai, India

  • 3. School of Built Environment, Engineering and Computing, Leeds Beckett University, Leeds, United Kingdom

  • 4. Department of Computer Science and Engineering, Chennai Institute of Technology, Chennai, India

  • 5. Centre for Research Impact & Outcome, Chitkara University Institute of Engineering and Technology, Chitkara University, Rajpura, Punjab, India

  • 6. Symbiosis Institute of Computer Studies and Research (SICSR), Symbiosis International (Deemed University), Pune, India

  • 7. Széchenyi István University, Győr, Hungary

  • 8. Department of Electrical and Electronics Engineering, The University of Manchester, Manchester, United Kingdom

Abstract

Introduction:

Cardiovascular health is increasingly at risk due to modern lifestyle factors such as obesity, smoking, stress, hypertension, and sedentary behavior. Post-pandemic health practices and medication side effects have further contributed to rising cases of early heart failure, particularly among individuals aged 25–40 years. This highlights the need for an automated and interpretable framework to predict heart disease at an early stage.

Methods:

In this study, body vitals acquired from a secondary dataset. Machine learning models including Support Vector Machine, Random Forest, Decision Tree, and Logistic Regression were employed for classification. Model performance was evaluated using accuracy, F1-score, and k-fold cross-validation.

Results:

Among the tested models, the Random Forest classifier demonstrated superior performance with an accuracy and F1-score of 0.955. The interpretability is enhanced with model predictions were explained using Local Interpretable Model-Agnostic Explanations (LIME) for local surrogates and SHAP values for global surrogates. SHAP decision plots provided clear insights into classification behaviour and feature contributions.

Discussion/Conclusion:

The proposed interpretable machine learning framework successfully predicts heart disease with high accuracy while maintaining transparency in decision-making. With the integration of sensor data with cloud-based analysis and explainable AI techniques, this study contributes to reducing the incidence of early heart failures and supports more reliable decision-making in healthcare applications.

1 Introduction

Cardiovascular disease (CVD) has always posed a serious threat to human beings and remains the primary cause of death globally. Heart diseases can cause substantial risk to the life of a person and significantly impact human health and wellbeing (). The World Health Organization (WHO) has reported that 18 million persons die of CVD every year, which represents 32% of all deaths worldwide, of which 85% are due to heart attacks. The WHO has stated that more than 70% of heart diseases occur in developing countries (). The World Heart Federation has predicted nearly 23 million CVD-related deaths by 2030, and the American Heart Association has reported that by 2035, nearly 130 million adults will contract heart diseases (). Heart diseases encompass various conditions such as irregular heartbeats, cardiomyopathy, arrhythmia, and peripheral or coronary artery that affect the heart and cause a global hazard health with serious medical manifestations (, ). The crucial risk factors for cardiovascular diseases are tobacco use, alcohol, unhealthy diet, physical inactivity, obesity, high blood pressure, diabetes, high cholesterol, and emotional implications (). WHO has been deliberately making efforts to curtail the encumbrance of heart diseases by implementing prevention and control efforts. Members of the WHO are planning to facilitate drug and counseling treatments for a minimum of 50% of people with a high risk of cardiovascular disease by the end of this year (). COVID-19 is an infectious pathogen that has created an aberrant impact on public health worldwide. It has been observed that COVID-19 infection has created an independent risk factor for heart diseases in some patients and has caused severe damage to the heart muscles, leading to myocarditis or heart failure. Blood clots and cardiac arrhythmias are the major risk factors for elevated mortality risks. Many emerging pieces of evidence and observational studies have been reported by researchers to the effect patients infected with the COVID-19 had suffered from impairment of myocardial function (, ) and cardiovascular complications (, ) such as myocarditis, arrhythmias, pericarditis, myocardial infarction, thromboembolism, stroke, and sudden death (, ). More than 72.3% of people, or over 5.55 billion individuals worldwide, have been administered a dose of COVID-19 vaccination to protect against virus variants effectively. However, vaccination intake rates have substantially stagnated for several reasons, and one among them is vaccine hesitancy (, ). The main reason for the reluctance on the part of people to get the vaccines administered is the side effects associated with cardiac complications like myocarditis and pericarditis (). The vaccine side effects are associated with a high risk of myocarditis that is highest in males between the ages of 16 and 24 years (). The studies conducted () show that males who received the second dose of the COVID-19 vaccine had the highest rate of cardiovascular complications (). The Center for Disease Control and Prevention (CDC) is continuously monitoring and conducting various surveys on patients having symptoms such as chest pain, palpitation, pounding heart, and shortness of breath and advising them to take the related medical tests for diagnosing myocarditis and pericarditis (). Monitoring patients with a high risk of heart disease is paramount to ensure their wellbeing and optimize treatment outcomes. Regular health monitoring helps clinicians to effectively assess the patient’s health condition, medical adherence, and drug intake adjustments. Health monitoring also helps patients understand their health conditions, disease progression, treatments, and self-care health management and prevents them from having the risk of adverse events ().

The Internet of Medical Things (IoMT) is a transformative and revolutionary technological concept in the field of healthcare used for amalgamating medical resources connected with network technologies for monitoring, predicting, and preventing health-related diseases (). The prognostic potential of the IoMT has fascinated the healthcare industry in terms of facilitating real-time surveillance using smart medical devices connected with software applications. The IoMT captures real-time data on patients using wearable devices, remote monitoring devices, connected medical equipment, implantable medical devices, mobile health applications, smart home medical devices, and point-of-care testing devices (). The IoMT devices play a vital role in health monitoring, data collection, personalized care, and transmission of real-time data for the decision-making process by healthcare providers. The communication system in the IoMT enables connectivity and data exchange among patients and healthcare providers for improving the patient’s health condition and enhancing overall healthcare management (). The IoMT is used in this work to monitor the patient’s health-related risk factors for cardiovascular diseases, such as tobacco use, use of alcohol, diet, physical activity, obesity, high blood pressure, diabetes, high cholesterol, and emotional implications for diagnosing, monitoring, and preventing heart diseases. Smart medical devices track the heart rate, electrocardiogram (ECG), heart rhythms, heart’s electrical activity, blood pressure, insulin level, sleeping level, physical activity, and stress management ().

The integration of IoMT technology into Artificial Intelligence has the potential to optimize the healthcare decision-making process by analyzing real-time data to develop predictive models (). The emerging utilization of AI and machine learning (ML) models has the potential to revolutionize healthcare management by enabling automation and analyzing data from IoMT devices for identifying symptoms and improving decision outcomes (). AI models predict disease progression, detect abnormalities and risk patterns, and facilitate interventions to avert adverse events (). However, AI approaches are often called “black boxes” due to a lack of interpretability and accountability. The high dimensionality feature of AI techniques makes it tedious for humans to interpret and understand the decisions taken (). Patients can face difficulty understanding the decisions and insights the machine learning models produce. Explainable AI (XAI) is considered a magic box to counteract the black box nature of AI models by providing optimal solutions with transparency in the healthcare industry (). The cutting-edge XAI technology is a game changer by as it generates explanations, visualizations, and justifications for the decision outcomes produced by AI models (). The healthcare industry can utilize this trailblazing XAI model to make clinical decisions transparently by allowing healthcare providers and patients to interpret the underlying reasoning behind AI’s decisions. An XAI model can be used for treatment recommendation plans that help physicians understand and comprehend appropriate interventions. This improves the trustworthiness among healthcare providers and patients for the successful implementation and deployment of AI-based healthcare systems ().

The purpose of this study is to develop an accurate and interpretable XAI framework to predict heart disease using input received through the IoMT sensors. This study also clearly understands the importance of medical parameters that provide transparency and consistency in the predictions. In this study, Section 2 describes the review of the literature on the prediction of heart disease on existing works. Section 3 discusses the materials and methods where a description of datasets, system architecture, and a mathematical model is added. Section 4 presents the results and Section 5 provides a discussion of the results. Finally, the paper concludes with a comprehensive conclusion section.

2 Background

The foundation for the enrichment of research is based on a comprehensive literature survey and relevant investigation in the respective domain. This section focuses on several research methodologies and literature reviews on heart disease prediction, IoMT-based health monitoring, machine learning models, and XAI. The current research methodologies and their respective strengths are identified along with their limitations. Heart diseases are considered a potential threat to human life and a leading reason for morbidity and mortality. The WHO has outlined the critical risk factors of heart diseases such as tobacco use, alcohol, unhealthy diet, physical inactivity, obesity, high blood pressure, diabetes, high cholesterol, and emotional implications depicted in Figure 1. A meta-analysis reveals various cardiovascular complications associated with COVID-19 and vaccinations. Figure 2 illustrates the cardiovascular disease complications of COVID-19 and notes the prevalence of myocardial injury, acute cardiac injury, arrhythmias, and heart failure, all of which elevates the risk of mortality.

Figure 1

Figure 2

2.1 Related works

Kumar et al. () analyzed the classification of heart diseases prediction, which involved five stages such as heart disease detection and diagnostics, machine learning models and algorithms used for healthcare, feature engineering and optimization techniques, evolving and advanced techniques in healthcare, and different applications of AI across various diseases and health conditions. This study analyzed the deep learning models for early diagnosis of heart disease prediction with evaluation techniques like sensitivity, specificity, and area under the curve (AUC). It also discusses ethical issues, dataset challenges, and transparency of the model. This paper clearly pointed out the advantages and challenges of modern equipment and advanced prediction systems. Rajkumar et al. () proposed IoT-based framework and advanced and enhanced deep learning framework. A Hungarian heart disease dataset is used, which is preprocessed by a median studentized residual approach for reducing error values and missing data. The Harris Hawk Optimization (HHO) approach is applied to select the features during preprocessing, which are classified using Modified Deep Long Short-Term Memory (MDLSTM). This output is updated by the Improved Spotted Hyena Optimization (ISHO) algorithm and achieved 98.01% accuracy in implementation. Hammadi et al. () presented an updated framework for cardiovascular disease prediction with the hybrid ensemble learning method. A soft voting method is introduced. Class imbalance is given preference here and the model achieves 97.4% accuracy in score 1, 83.6% accuracy in score 2, and 93% accuracy in score 3. This advanced technique applies ensemble learning for early detection of heart diseases. The Department of Computer Science & Engineering BRAC University in Dhaka, Bangladesh. Rokoni et al. () focused on model interpretability and applied one-dimensional Convolutional Neural Networks (1D CNNs) and logistic regression for classifying diseases. It achieves 80% overall accuracy, and Local Interpretable Model-Agnostic Explanations (LIME) provides transparency by finding the influenced features like glucose, blood pressure, and troponin.

Wang and Song () presented an edge-assisted IoMT framework for monitoring aged people having chronic diseases. The IoMT-based smart home monitoring model is utilized to access medical data and diagnose diseases for aged persons by continuously monitoring and communicating hastily using edge computing. Martinek et al. () used federated learning and blockchain technology to ensure privacy for healthcare monitoring. The IoMT technology used in this work employs sensors for health monitoring, and the sensed data are stored and managed using a fog-cloud-assisted network. The federated learning and fraud detection mechanism–enabled blockchain framework is used to process application workloads and validate the quality of service. Kumar et al. () proposed a novel IoMT-based healthcare monitoring system using rooted elliptic curve cryptography with Vigenere cipher (RECC-VC) for securing the environment. RECC-VCC enhances security, and the exponential anonymity model is used for privacy protection. The Improved Extension Neural Network (IENN) framework is used to analyze the level of sensitive data, and the Gaussian mutated chimp optimizer is used to update the weight. Blockchain technology is used to store and manage transactions on the cloud server. Kishor and Chakraborty () presented an IoT-based health monitoring system by predicting various diseases such as heart disease, diabetes, breast cancer, dermatology, thyroid, liver disease, and surgery data using machine learning approaches like Decision Tree, Naïve Bayes, Random Forest, Support Vector Machine (SVM), Adaptive Boosting, Artificial Neural Network, and K-nearest neighborhood (KNN). Shafiq et al. () presented a deep learning framework using CNN for detecting heart disease symptoms by analyzing the biosensor input for detecting heart disorders. The PASCAL dataset is used to train the CNN model, and the real-time data sensed by IoT sensors are stored in the cloud. The sound of the heart is given as input for classifying whether patients are affected by heart disease. Kumar and Gandhi () introduced a three-tier IoT architecture for detecting heart disease. Wearable sensors are used to observe the patient’s health condition, and Apache HBase stores the patient’s monitored data in cloud computing. Apache Mahout is utilized to implement a logistic regression framework for heart diseases. Panja et al. () utilized IoMT architecture to monitor and assess the health issues of infected patients and trigger an alert message to their clinicians and relatives. The real-time data collected from patients are transmitted to the cloud via edge devices for processing. A severity analysis of the infection is carried out using fuzzy logic to detect the risk status of COVID-19 patients effectively. Jain et al. () proposed point-of-care testing to rapidly detect infectious diseases and give spot results for taking early action.

IoMT devices are also used to capture the patient’s vitals for early detection of diseases such as malaria, influenza A, Ebola virus, Zika virus, COVID-19 virus, and dengue fever. Rezaee et al. () suggested a meta-heuristic fuzzy inference system for emotion recognition using the IoMT. Patients are tested by playing music videos to detect their emotional states. Electroencephalography (EEG) signals are captured before and after meditation. Using an optimized, innovative Gunner algorithm, a fuzzy inference-based classification approach is used to classify emotions. Krupa et al. () presented an IoMT-based deep learning framework for automatically detecting fetal QRS. The framework used two methods: one for detecting fetal QRS complex using a deep neural network and the second for classifying the results by acclimatizing transfer learning to improve accuracy. The method uses a time-frequency image as input for an IoT-based deep neural network in the abdominal ECG without removing the maternal components. Lu et al. () developed a novel IoMT-based fetal monitoring model incorporating an automatic fetal heart rate (FHR) rating method to evaluate fetal health conditions inside the uterus using digital cardiotocographic signals. The monitoring system uses Kreb’s Fischer, improved Fischer, and American College of Obstetricians and Gynecologists (ACOG) classifiers to detect and classify fetal conditions as good and bad for comparison. Rahmani et al. () used fog computing–based e-Health gateways by offering higher-level services for storage provisioning and data processing to form a geo-distributed middle layer between the cloud and IoT sensors. The framework uses an early warning score for monitoring health to facilitate energy efficiency, interoperability, reliability, mobility, and performance. Nandy et al. () proposed an IoMT-based intelligent agent mechanism to detect brain response using an electroencephalography signal. A bag of neural network categorizes the complex brain signals captured by the IoMT sensors and detects the brain responses. The IBoNN framework is compared with standard machine learning algorithms. Yadav et al. () presented biomarker-based electrochemical immuno sensors for diagnosing COVID-19 using the IoMT and artificial intelligence. The smart sensing technique is used with bioinformatics approaches for monitoring non-invasive SARS-COV2.

Verma et al. () summarized nano-integrated wearable biosensors and the use of 5G in the Internet of Things for healthcare applications. Fouad et al. (54) presented a numerical approach using the Gautschi model for vertebral tumor prediction. The IoMT technology is used for predicting tumors employing heuristic hock transformation for evaluating possible perpetual incapacity caused by tumors on Haar-like characteristics (HLC), logistics models (LM), conservative therapy method (CTM), and carbon fiber including reinforced materials (CFRM) approaches. Figure 3 describes the percentage of AI and machine learning approaches used in the healthcare industry. In recent years, machine learning models have revolutionized the diagnosis of various diseases and the assessment of the risk factors involved in making accurate decisions. According to a background study, supervised learning approaches like 35% of logistic regression, 26% of decision tree, and 24% of neural networks, as well as 5% of unsupervised learning methods like clustering and anomaly detection, have been used by the healthcare industry to assess the risks.

Figure 3

Ribeiro et al. (55) presented a novel agnostic approach called LIME, which is used to provision the comprehensiveness of the decisions made by the banking staff by determining them with simple and interpretable models. The LIME model helps improve accuracy, transparency, and trustworthiness. This model can be used with complex machine learning algorithms without any knowledge of their working mechanisms. Lundberg and Lee (56) proposed a new technique, SHAPELY Additive exPlanations (SHAP), to interpret the existing complex machine learning models. The SHAP algorithm provides global and local interpretations to help borrowers easily understand the predictions made by black box techniques. SHAP can be used in various machine learning approaches and deep neural networks and effectively work with real-world datasets.

Muddamsetty et al. (57) presented an evaluation for understanding the outcomes of machine learning models. Thus, it is evident that the XAI model helps to present outcomes with transparency and provides solutions to black box models. Explanations for the clinical prediction outcome entail the justification of reliability and trustworthiness that can be achieved using XAI models (58). Onan (59) presented a hierarchal graph-based model for text classification of dynamic fusion using BERT. The framework uses seven stages for graph-based text classification and analysis with various benchmark functions. Onan (60) proposed a genetic technique combined with graph-based neural networks for generating augmented text having high dimensional feature space. The objective function is based on perplexity when evaluating the quality of generating augmented text data. Onan (61) proposed a Semantic Role Labeling algorithm with an Ant colony optimization approach for generating training data to improve the performance of the natural language processing (NLP) framework. The semantic roles are identified using semantic role labeling (SRL) for text augmentation to enhance the quality of training data. Onan (62) suggested a bidirectional convolutional recurrent neural network framework for semantic analysis using gated recurrent unit (GRU) and LSTM layers. Feature extraction is carried out by the bidirectional layers to reduce dimensionality and extract high-quality features. Onan et al. (63) presented a two-stage topic extract model using a word embedding approach and cluster analysis. The word vectors are extracted by Word2Vec, POS2vec, LDA2vec, and word position2vec schemes. A comparison of Naïve Bayes, SVM, Random Forest, and Logistic regression with ensemble methods is used for evaluating the statistical key extraction model (64). Onan et al. (65) presented a consensus cluster mechanism using an undersampling model with five supervised learning algorithms and three ensemble learners for imbalanced learning. Onan and Korukoğlu (66) utilized an ensemble model for feature selection with a genetic-optimized algorithm for sentimental analysis. Sentiment classification based on a hybrid ensemble pruning model with consensus clustering is utilized for text classification (67). Sentimental analysis for product reviews (68), online course evaluation (69), and mining opinions for instructors (70) is done using deep neural networks. Onan (71) presented a comparative analysis of feature engineering models using five base learners for text genre classification and language function analysis. Onan and Toçoğlu (72) suggested inverse gravity moment utilizing bidirectional LSTM for representing text documents. The LSTM framework is evaluated on the basis of the sarcasm identification corpus. The deep learning model is utilized to identify sarcasm for predicting the performance of sentiment analysis. Vakharia et al. (73) proposed three deep learning frameworks with optimized explainable artificial intelligence for predicting the discharge capacity of the battery. The jellyfish optimization algorithm is used with the XAI model to improve the predictive performance. Ali et al. (74) presented an SVM model based on four Ant Bee Colony (ABC) algorithms, Genetic Algorithm (GA), Particle Swarm Optimization (PSO), and Whale Optimization Algorithm (WAO) for tuning hyperparameters. A teaching learning–based optimization algorithm with a heat transfer searching model is used to select features and identify faults. Suthar et al. (75) and Vakharia et al. (76) highlighted a comparative study of feature ranking approaches for fault identification by using the Fisher score, ReliefF, Gain ratio, Wilcoxon rank, and Memetic feature selection model. The literature survey shows that the XAI framework improves prediction accuracy with interpretability and explainability in healthcare applications because of the crucial nature of the decision-making process and public health safety. Tables 1, 2 depict a comparison of various black box models used for healthcare applications.

Table 1

ReferenceTitleAdvantagesResearch gap
Hashem et al. (77)Predicting neurological disorders linked to oral cavity manifestations using IoMT-based optimized neural networks
  • Reduced complexity

  • Oral cavity linked nervous problem detection rate

  • Minimized the feature dimension

  • Interpretability vs. accuracy trade-off

  • Limited explanation of complex models

  • Scalability

Zhu et al. (78)IoMT-enabled real-time blood glucose prediction with deep learning and edge computing
  • The wearable sensor’s power and memory footprint are analyzed

  • Prediction accuracy for three datasets

  • Scalability.

  • Limited applicability

  • Interpretability vs. accuracy trade-off

Abbas et al. (79)Secure IoMT for disease prediction empowered with transfer learning in healthcare 5.0, the concept and case study
  • Model performance

  • Generalizability

  • Security

  • Limited explanation of complex models

  • Scalability

Nandy et al. (80)An intrusion detection mechanism for a secure IoMT framework based on swarm-neural network
  • Security. High performance due to optimization

  • Limited applicability

  • Interpretability vs. accuracy trade-off

  • Scalability

Lakhan et al. (81)Federated learning–based privacy preservation and a fraud-enabled blockchain IoMT system for healthcare
  • Privacy preservation

  • Minimum energy consumption

  • Limited number of models

  • Lack of external validation

  • Limited scope of interpretability

Wang and Song ()An edge-assisted IoMT-based smart-home monitoring system for the elderly with chronic diseases
  • Local medical data diagnosis and rapid communication

  • Scalability

  • Limited applicability

  • Interpretability vs. accuracy trade-off

Zhang et al. (82)A joint deep learning and internet of medical things–driven framework for elderly patients
  • Energy efficiency

  • Sustainability

  • Reliability during data transmission

  • Limited to a specific set of datasets

  • Does not incorporate interpretability techniques

  • Does not incorporate explainability techniques

Khan and Algarni (83)A healthcare monitoring system for the diagnosis of heart disease in the IoMT cloud environment using MSSO-ANFIS
  • Better accuracy

  • Improved convergence rate

  • Lack of external validation

  • Limited scope of interpretability

Guleria et al. (84)XAI framework for cardiovascular disease prediction using classification techniques
  • Comprehensive evaluation,

  • Large dataset

  • Transparent evaluation criteria

  • Lack of external validation

  • Limited scope of interpretability

  • Limited to a specific set of datasets

Motivation for the proposed work from the review perspective.

Table 2

ReferenceAlgorithms comparedType of dataPrediction performance
Juhola et al. (85)ANN, NBDisease symptomAccuracy: (ANN=85, NB=88)
Long et al. (86)ANN, LRClinical and demographic dataAccuracy: (ANN=0.965, LR=0.963)
Palaniappan and Awang (87)ANN, DT, SVMClinical data for cancer incidence and survivalAccuracy: (ANN=0.947, DT=0.936, SVM=0.957)
Jin et al. (88)LR, RFElectronic health recordsAccuracy: (LR=0.663, RF=0.627)
Puyalnithi and Viswanatham (89)DT, NB, RF, SVMClinical and demographic dataSensitivity: (ANN=0.956, DT=0.958, SVM=0.971)
Forssen et al. (90)LR, RFMetabolomic dataAccuracy: (LR=0.767, RF=0.732)
Tang et al. (91)ANN, LRClinical, demographic, behavioral, and medical dataSpecificity: (ANN=0.928, DT=0.907, SVM=0.945)
Toshniwal et al. (92)ANN, LRClinical and demographic dataAccuracy: (ANN=0.909, LR=0.897)
Yang et al. (93)ANN, DT, LRClinical and demographic dataAccuracy: (ANN=0.909, DT=0.935, LR=0.894)
Mustaqeem et al. (94)DT, RF, SVMImage dataAccuracy: (DT=0.932, RF=0.963, SVM=0.959)
Mansoor et al. (95)DT, KNN, NBElectronic health records, medical image, and gene dataAccuracy: (DT=0.646, KNN=0.454, NB=0.495)
Kim et al. (96)LR, NB, SVMGut microbiotaAccuracy: (LR=0.98, NB=0.94, SVM=0.99)
Taslimitehrani et al. (97)ANN, LR, SVMElectrochemical measurements of salivaAccuracy: (ANN=80.70, LR=75.86, SVM=84.09)
Anbarasi et al. (98)DT, NBClinical and demographic dataAccuracy: (DT=99.2%, NB=96.5%)
Bhatla and Jyoti (99)ANN, DT, NBClinical dataF1-score: (ANN=80.20, LR=75.71, SVM=84.06)
Thenmozhi and Deepika (100)KNN, LR, SVMDemographic, anthropometric, vital signs, diagnostic, and clinical laboratory measurement dataAccuracy: (KNN=79.5, LR=80.7, SVM=82.6)
Tamilarasi and Porkodi (101)KNN, LR, NB, RF, SVMDemographic and clinical test resultAccuracy: (KNN=0.721, LR=0.755, NB=0.762, RF=0.803, SVM=0.749)
Marikani and Shyamala (102)ANN, LR, RF, SVMDemographic, anthropometric, diagnostic and clinical lab measurement dataAccuracy: (ANN=0.931, LR=0.935, RF=0.930, SVM=0.986)
Lu et al. (103)ANN, NB, SVMClinical, demographic, and diagnostic dataAccuracy: (ANN=86.04, NB=82.31, SVM=86.62)

Comparison of algorithms and prediction performance.

ANN, artificial neural network; NB, Naïve Bayes; LR, logistic regression; DT, decision tree; RF, random forest.

2.2 Research questions

This study aims to answer all the following questions:

  • 1.

    How can this IoMT sensor device help predict the heart disease effectively?

  • 2.

    Which type of machine learning or deep learning methodologies are most relevant for predicting heart disease accurately?

  • 3.

    How can feature selection techniques help identify the most dominant parameters for prediction?

  • 4.

    What are all the challenges in integrating the IoMT with real-time heart disease prediction and how can they be addressed?

  • 5.

    Does the proposed framework apply effectively in real-time analysis problems for early heart disease prediction?

This work addresses the drawbacks of state-of-the-art ML-based techniques in achieving increased transparency, interpretability, and accountability with high-accuracy outcomes.

2.3 Feature selection and existing work

The IoMT integrates wearable devices with sensor technologies to monitor health parameters, track physical activity, and enable remote patient monitoring. The literature survey focuses on two specific sensor categories: vital signs and motion sensors. For vital signs sensors, Rao et al. (104) present a non-invasive wearable device that accurately monitors BP without requiring invasive catheterization. The device utilizes capacitive wrist and/or foot sensors to acquire pulse waveform data, which are then processed using artificial neural networks to determine systolic, diastolic, and mean arterial pressures. A comparison with invasive arterial line data confirmed the device’s accuracy, making it a viable alternative for continuous BP monitoring in critically ill infants. In motion sensors, Jakob et al. (105) evaluate the effectiveness of wearable sensors in analyzing motion patterns in individuals with Parkinson’s disease. The study assesses the accuracy and reliability of the sensor system in detecting and quantifying motor symptoms associated with Parkinson’s disease, such as bradykinesia and shuffling gait. Wearable sensors distinguish Parkinson’s patients from healthy controls, showing their potential for clinically relevant gait assessments in flexible environments. These research papers are examples of studies conducted on sensors used in IoMT projects. The surveyed literature demonstrates the significance of vital signs sensors in non-invasive blood pressure monitoring and the potential of motion sensors in analyzing motor symptoms in Parkinson’s disease. The following parameters mentioned in Figure 4 are taken into consideration while designing existing and ongoing IoMT systems that are positively helping to transform the IoMT domain through cutting-edge technology. Physiological parameters, biochemical parameters, electrical activity, respiratory parameters, motion and activity, sleep patterns, environmental factors, and medical adherence are the IoT devices applied for treatments. Some of the IoMT projects have incorporated the above-mentioned parameters and have made the availability of diagnostics more accessible, for example, Ovularing (106), VitalPatch (107), SmartPill (108), GlucoWear (109), MindMotion Pro (110), BioStampRC (111), Biotricity (112), SmartMat (113), PillCam (114), WAND (Wireless Artifact-free Neuromodulation Device) (115), Abilify MyCite (116), Embrace (117), and Insulet Omnipod (118).

Figure 4

3 Materials and methods

In this work, we propose an IoMT-based heart disease prediction framework based on machine learning models like Logistic Regression (119), SVM (120), Decision Tree (121), Gradient Boost (122), and Random Forest (123). Figure 5 depicts the layered architecture of the work having four layers of IoMT: a device layer, cloud layer, machine learning models layer, and Explainable AI layer. The IoMT devices capture patients’ vitals, and the sensed patient’s data are transferred to cloud storage. The machine learning models herewith are used to detect and classify cardiovascular diseases and associated risk factors for diagnosing, monitoring, and preventing heart diseases. XAI techniques such as LIME (124) and SHAP (125) help overcome the limitations of traditional Machine Learning models by providing interpretable decision outcomes, thereby assisting both patients and clinicians.

Figure 5

3.1 Importance of XAI in the IoMT

The following case studies outline why XAI will prove to be a revolutionary change required in the IoMT.

3.1.1 Case study 1: a 26-year-old adult died due to cardiac arrest

A 26-year-old man collapsed suddenly at a Metro Station in New Delhi because of cardiac arrest. The young man was immediately taken to a hospital, and the physician declared that the person died because of chronic fat deposits in the arteries. The postmortem was carried out at a medical institute, which revealed that the visceral organs and the brain were congested, resulting in lung blockage. Following this incident, the healthcare administration raised concerns regarding the prevalence of abrupt cardiac deaths among young adults and drew attention to the presence of undiagnosed cardiovascular diseases. Clinicians are advised to consider the risk factors and causes of heart disease and to take preventive measures for early diagnosis and further treatments.

3.1.2 Case study 2: a 40-year-old actor’s demise due to massive cardiac arrest

A 40-year-old man and actor died because of a massive heart attack at his Mumbai residence. He took medicine, slept, and did not wake up. He was immediately taken to the Cooper Hospital in Mumbai, and the clinicians declared that the person was brought dead due to a massive heart attack. Clinicians worldwide are advised to assess the risk factors and lifestyle changes, thereby stressing regular health checkups that can help prevent cardiovascular diseases.

Table 3 provides an overview of the sensors described above and elucidates the shortcomings and advantages of these devices, which acted as a support to this work.

Table 3

Sl. no.TypeApplicationParametersAdvantagesMeritsDemerits
1Ovularing (106)WearableWomen’s healthOvulation monitoringAccurate fertility trackingLimited compatibility with other devices
2VitalPatch (107)WearableHealthcareVital signs monitoringReal-time health monitoringRequires regular battery replacement
3SmartPill (108)IngestibleHealthcareDrug delivery monitoringNon-invasive medication trackingPossibility of device malfunction
4GlucoWear (109)WearableDiabetes careContinuous glucose monitoringImproved glucose managementCalibration requirements for accuracy
5Bio Stamp RC (110)WearableResearchMotion analysisLong-term data collectionLimited sensor placement options
6Biotricity (111)WearableCardiologyECG monitoringReal-time cardiac monitoringRelatively high cost for consumer use
7Mind motion pro (112)Bio-feedback devicesVarious rehabilitation applicationsMuscle activity, EMGProvides real-time feedback for muscle controlRelies on accurate sensor placement and signal quality
8SmartMat (113)WearableFitnessYoga and exercise trackingPrecise posture and movement analysisLimited battery life
9PillCam (114)IngestibleMedical imagingGastrointestinal imagingNon-invasive imaging of the digestive systemLimited imaging capabilities compared with MRI
10WAND (115)ImplantableNeurologyDeep brain stimulationEffective treatment for neurological disordersInvasive surgical procedure for implantation
11Abilify MyCite (116)IngestibleMental healthMedication adherenceMonitors medication ingestionLimited availability and regulatory approval
12Empatica Embrace (117)WearableEpilepsySeizure detectionAlerts caregivers during seizuresSome false alarms and limitations in accuracy
13Insulet Omnipod (118)WearableDiabetes careInsulin deliveryTubeless insulin pump systemInitial setup and learning curve for users

Comparison of wearable and ingestible health devices.

3.2 Dataset description

CVD takes the lives of around 18 million people every year and is the primary cause of death. The rate of accountability of death reports due to CVD is around 31%. A total of 80% of deaths associated with CVD are mainly due to heart attack and stroke. These attacks are observed in groups of people who are less than 70 years old. With this in mind, a dataset (126) has been prepared as an amalgamation of observations recorded from Cleveland (303), Hungarian (294), Switzerland (123), Long Beach, VA (200), and Stalog Dataset (270). After removing duplicates, the final dataset contains 918 instances with 11 important features for analyzing CVD diseases. The dependent target class is Heart Failure. The other independent features are Age, Sex, Chest Pain Type, ST_Slope, Cholesterol, Resting BP, Blood Sugar, Resting ECG, Exercise Angina, Old Peak, and Maximum Heart Rate (MaxHR). Some of the features are numeric, and some of the features are non-numeric. Table 4 provides the list of features converted to numeric data. The string data are transformed using the Label Encoder preprocessing technique with Min-Max scalar transformation. The dataset (126) has no missing values or class imbalance. Heart Rate, Variable Heart Rate, Blood Glucose Level, MaxHR, Blood Pressure, and ECG (Polar H10 Sensor) are measured by IoMT sensors. The other readings are observed in the oscilloscopes and treadmills (ST_Slope, Restring Angina), and some data are collected directly from patients and their attenders (Name, Age, Sex, etc.).

Table 4

Sl. noFeaturesTypesNumeric changeTransformation
1Chest pain typeATA1Label encoder and min-max scalar
NYP2
ASA3
TA4
2ST_SlopeUp1Label encoder and min-max scalar
Down0
Down zero1
3Resting ECGNormal0Label encoder and min-max scalar
Abnormal1
4SexMale1Label encoder and min-max scalar
Female0
5Exercise anginaYes1Label encoder and min-max scalar
No0

Feature conversation details of the dataset.

NYP, non-anginal pain; TA, typical angina.

3.3 System architecture

Figure 6 provides an overview of interfacing the ML algorithms discussed in this section with XAI. In terms of monitoring and managing cardiovascular health, IoMT devices play an important role in the use of advanced transformation techniques. These devices increase power connectivity, analyze data, and monitor remotely, providing advanced care for cardiac patients and improving patient health. There are many IoMT devices for monitoring the heart behavior of patients, such as remote ECG monitors, wearable heart rate monitors, pacemakers, BP monitors, temperature monitors, and medication dispensers. These devices help healthcare professionals to monitor patients continuously. They can personalize treatment plans, and they can easily predict previous symptoms and take immediate action.

Figure 6

Machine learning techniques play a substantial role in identifying heart diseases with the help of IoMT devices. Data are collected from IoMT sensor devices, and ML algorithms understand the data, detect anomalies, and produce solutions for accurate heart diagnosis. Many ML algorithms can be applied to train the model to analyze data and recognize patterns. A large volume of data can be processed by ML algorithms from IoMT sensor devices, such as blood pressure measurements, ECG reading results, and heart rate information. ML algorithms like decision trees, Random Forest, SVM, and Logistic Regression are applied here with IoMT devices.

AI algorithms suggest transparent and interpretable explanations for making decisions or predictions. Traditional AI algorithms work as black boxes and produce results with less transparency. When we use explainable AI, it produces an understanding of the reasoning behind its results. Data are collected from IoMT devices and sent for preprocessing, followed by model selection and training. After training the data, feature analysis is done, and during local interpretability, explainable AI uses LIME to understand how specific features contribute to identifying the heart disease. As global interpretability, SHAPLEY helps explainable AI analyze overall behavior and features and their relationship to the decision-making process. Model-agnostic interpretability independently understands the prediction process and aims to apply it to any algorithm. The results can finally be visualized so that appropriate decisions can be made.

3.4 Mathematical modeling

3.4.1 Random Forest

Random Forest (127) is an ensemble technique of a machine learning algorithm applied for classification and regression problems. The ensemble combines many models to make predictions accurately. To make an accurate prediction, a Random Forest combines many decision trees (128). Forest refers to a collection of decision trees. Every tree is made independently by a subset of the training data and its input features. Selecting data and features randomly reduces the overfitting problem and creates diversity among each tree. Random forest considers the majority vote from different samples of the decision trees for classification and regression tasks. Bagging or bootstrap and boosting are the two types of ensemble methods. Bagging depends on majority voting by creating many training subsets from the training sample with replacements. Boosting refers to joining weak and strong data by making sequential models to produce the highest accuracy. When the amount of data in the training set is , then with replacement “,” data are sampled at random as a bootstrap sample. This helps to grow the tree with training data. When there are “” input variables, is chosen so that “” variables are taken at random from “.” The value “” is constant when the tree grows to the maximum extent. Many subtrees made by the parameters are formed in the forest. When the forest is completely trained for classification, it is traversed across all the subtrees (129). The classification result from each tree is taken as a vote. The maximum vote is considered a new instance. The generalization error () for the Random Forest is given by Equation 1.Here, is a margin function that measures the average number of votes from exceeding any other class. refers to the prediction variable and refers to the classification task. “” denotes the indicator function. The expected value for the margin function of a random forest is indicated as Equation 2.A Random Forest’s average strength and the base classifiers’ mean correlation are joined as generalization errors. If represents the mean rate of correlation, the generalization error value for the upper bound is given by Equation 3.To achieve better accuracy in a Random Forest, the subtrees of decision trees must be consistent and diverse. Random Forest is very efficient in detecting outliers. It is scalable, robust, and handles missing data without imputation.

3.4.2 Local Interpretable Model-Agnostic Explanations

LIME (130) is a post-hoc model-agnostic framework for any black box machine learning model’s judgment for all instances (55). LIME creates new data from the nearest neighborhood and finds the predictions of these new samples with the help of a black box model. LIME’s explanation depends on monitoring the classifier model’s behavior based on local surrogate models. The LIME algorithm follows three steps to train a surrogate model.

  • 1.

    Select a few data instances as , representing the reason for an opaque recommender model predicting the feature vector for the probability . LIME expects the data to be converted into an interpretable picture like a binary vector representing the available/non-available components.

  • 2.

    Create a new dataset of perturbed data by taking non-zero elements of at random. The labels must be identified for this new set of data elements in in the closest area of . To obtain the labels for the new data, the perturbed samples are transformed back into the original form . The opaque model is then examined for each instance . Because the perturbed samples are randomly generated, there might be samples that are closer or farther away from the original instance for weighing. This weight is measured as to evaluate the closeness between the data and .

  • 3.

    Using this newly weighted data and the labels created by , a new model is trained, where refers to models such as decision trees, linear models, and so on. The interpretable and explanatory surrogate model of the new data is then used to explain as shown in Equation 4.Here, is the loss function, which measures how follows the behavior of in the nearest neighborhood of . Minimizing this loss function ensures that the behavior of aligns with the behavior of indicated by . The complexity of the model must be kept low. When is represented as a linear function, , the equation 5 becomes a linear regression problem to evaluate and .The advantages of LIME are that it is easy to implement, completely fast in terms of computational techniques, and easy to work with in tabular data, text, and images.

3.4.3 SHAPELY Additive exPlanations

The SHAP (124) method improves computational time, and tree-based methods improve explanation precision. The main goal of SHAP is to form perturbations to simulate the features that are not present and to use the linear local model to approximate the prediction changes as given in LIME. It ignores retraining the model without the feature of interest. Local explanations can be combined to describe the model’s global performance. Local and global explanations are reliable with each other as they follow the same basic methods. SHAP uses agnostic explainer KernelSHAP and model-specific explainers such as TreeSHAP for tree-based models, DeepSHAP for deep models, and LinearSHAP for linear models.

SHAP produces SHAPELY values, which express model predictions as linear combinations of binary variables. This framework explains how each covariate contributes when fixed in the model. The prediction , using , for a linear model for the binary values with the elements , is given by Equation 6.Here, is a variable for explanations which is shown in Equation 7.where is the model of this method, is the variable, and are the selected variables. The value denotes, for every prediction, the SHAPELY values from its mean value of the th variable. Local accuracy results from the explainable model are equal to those of the basic models. The missing nature of the SHAPELY values has features that were not added as the first input without any effect. Consistency of the model changes with reliance on a single feature, and related characteristics cannot be reduced independently of other factors. The advantage of SHAP is that it predicts an instance disseminated among the feature values. The limitations are its slow computational time, high computational complexity, and problems with explanation instability similar to LIME.

3.5 Algorithm

This section describes two algorithms, one for the heart risk evaluation through Algorithm 1 and the other for explaining heart failure through the Algorithm 2. These two algorithms comprehensively analyze and explain the risk of Heart Failure as a complete solution. In Algorithm 1, the performance metrics such as accuracy, precision, recall, sensitivity, specificity, and F1-score are evaluated. During the testing phase, when becomes 1, the heart failure alarm will be activated. Otherwise, the result indicates that the function of the heart is normal. In Algorithm 2, the model with local surrogates explains the appropriate decision after heart failure when the prediction is local. In case the probability of the prediction is global, explainability is achieved in global surrogates.

Algorithm 1

Input: ;
;
;
;
;
;
;
;
;
;
Accuracy: ;
Precision: ;
Recall: ;
F1-score: ;
Activation: ;
whiledo
if is Potable do
;
;
;
;
;
;
end
else
is Not Potable
;
;
;
;
;
;

Algorithm for heart disease prediction.

Algorithm 2

Input: ;
;
;
;
;
;
;
;
;
;
;
whiledo
if is local then
;
;
Decision Explained with Local Surrogates (LIME)
else
is global
;
;
Decision Explained with Global Surrogates (SHAPELY)
end

Algorithm for explainable AI.

3.6 Environment-based attribute access control algorithm

The dataset under consideration must be protected and authenticated. Hence, rigorous data access control permissions must be set in the cloud to access it properly. A secure environment-based attribute access control system is required in this context to protect unauthorized access to the data in the cloud. The model is divided into two categories: static and dynamic. Users with the lowest role, such as those looking for recommendations, access information in a static environment. This audience will only be permitted to obtain legal information; no other transactions will be permitted. In a dynamic state, different parameters are measured and recorded at various instances of time. Thus, many data acquisition and update cycles are a series of transactions carried out in the cloud in big time. These states only allow special users such as clinicians and administrators.

The development of a digital identity is the first step. The key used in the digital identity protects and guarantees a transmission between the server and the client, and the key is . The user shares its digital account identity and the symmetric key , and the corresponding data information can be obtained by decrypting the key . Various functions used in the algorithm, such as IssueRole, revokeIssueRole, and partialExtension, help the framework achieve a secured space to function. After the digital identity is authenticated and a role is identified, the model can access the framework accordingly. Each of the entity’s transactions is considered along with its authorization. Therefore, a secure environment for the fuzzy framework is achieved.

4 Results

4.1 Experimental setup

The 11 parameters that determine the failure of the heart are acquired from various sources across various countries and used in this work. These parameters have a strong influence on determining heart failure in real time. Most of these parameters are embedded with IoMT sensors, which can be integrated through information fusion in cloud platforms. Later, these data are classified by cloud machine learning models and transformed into a valid dataset. One such dataset is used in this work for experimental analysis. Because the problem is binary, the experimentation is done with machine learning models such as SVM, Logistic Regression, Decision Tree, and Random Forest. The explanation of this dataset is provided by LIME and SHAPELY values. The classification probability of the random forest model is evaluated due to its high classification accuracy with various explanations for clarity.

4.2 Results

4.2.1 Preprocessing

The dataset is preprocessed to convert the data types into a unified format, which makes it suitable for the classification problem. The statistical analysis of the various features of interest is tested with the correlation matrix shown in Figure 7. The features that have a higher correlation as per the correlation map are Exercise-Induced Angina, Chest Pain Type, and Age. Preprocessing equations

Figure 7

  • 1.

    Missing value imputation

    Let be a feature vector with missing entries.where is the mean of observed values and is the number of non-missing entries. Mean imputation replaces missing values with the average of the available values in the feature.

  • 2.

    Label encoding

    Let a categorical variable be transformed into integer labels as Equation 12.Each distinct category is mapped to a unique integer . This method is commonly used when categories have no intrinsic ordering.

  • 3.

    Standardization of binary target class

    Given a binary target variable , standardization is defined as Equation 13.whereAssuming , the mean and standard deviation are computed to transform into a zero-mean, unit-variance variable suitable for certain learning models.

4.2.2 Machine learning models

The target attribute Heart Disease is a binary classifier where “1” indicates heart failure and “0” indicates no failure. Because the problem is binary, we apply machine learning models such as SVM, Logistic Regression, Decision Tree, Random Forest, AdaBoost, and Gradient Boosting Classifier Algorithm. Model parameters and specifications of various methods are specified in Table 5. The model parameters of the Random Forest are slightly higher than that of the other models with respect to AUC. The results obtained in this work have only a thin difference in the metric values measured across various machine learning models since the dataset is free from missing values or class imbalance. The cost function of Logistic Regression (Equation 8), Gradient Boost (Equation 9) AdaBoost (Equation 10) are highlighted. The metric evaluation is presented in Table 6. There are essential metrics such as sensitivity and specificity, which estimate the true positive rate , true negative rate , false positive rate , and false negative rate . These parameters calculate the reliability of the model. The explanation of a machine learning model is based on reliability and performance. Table 7 presents these metrics with corresponding values for each machine learning model.

Table 5

ModelHyperparametersTime complexityCost function
Logistic Regression (131)Solver, penalty (optional)2–3 s
SVM (132)C Gamma Kernel size2–3 s
Decision Tree (133)Gini, max depth, minSamples, features2–3 s
  • Find best split in all variables that maximize impurity decrease

  • Label the with the best-split variable and its value

  • Divide the available learning data into and

  • Create nodes and that contain data and , respectively

  • Repeat with and data

  • Repeat with and data

Random Forest (127)max_depth Min_sample_split Max_leaf_nodes Min_samples_leaf N_estimators Max_sample (bootstrap sample) Max_features2–3 s
  • There are number of trees instead of only one tree

  • There are number of variables in each tree instead of , where and is the total number of variables

  • Each tree is built using number of samples, where is 63.2% of the total number of samples

Gradient Boost (134)Maximum iterations Learning rate Maximum depth Or maximum leaf nodes2–3 s
AdaBoost (135)Number of estimations Learning rate2–3 s

Model parameters and specifications.

Table 6

MethodAccuracyPrecisionRecallF1-scoreMCCROC
SVM (132)0.890.890.890.890.7770.94
Logistic Regression (131)0.8750.8750.8750.8740.7460.933
Decision Tree (133)0.9610.9620.9610.9610.9220.991
Random Forest (127)0.9550.9550.9550.9550.9100.994
Gradient Boost (134)0.9350.9350.9350.9350.8680.985

Classification report of the various machine learning models.

Table 7

MethodSensitivitySpecificity
SVM (132)0.890.89
Logistic Regression (131)0.8750.875
Decision Tree (133)0.9620.961
Random Forest (127)0.9550.955
Gradient Boost (134)0.9350.935

Sensitivity and specificity analysis of the various machine learning models.

4.2.3 Tenfold classification

Table 8 illustrates the results from 10-fold validation without preprocessing using a Python IDE. Without the application of preprocessing, the results provide accuracy, which is comparatively less than the original 70-30 train-test evaluation. The model has already been optimized with the highest levels of accuracy through preprocessing techniques. The preprocessed values are already tabulated in Table 6.

Table 8

MethodAccuracyPrecisionRecallF1-scoreAUC
SVM (132)0.8440.8440.8440.8440.904
Logistic Regression (131)0.8610.8600.8600.8610.924
Decision Tree (133)0.7920.7920.7940.7920.778
Random Forest (127)0.8590.8590.8590.8590.920
AdaBoost (135)0.7810.7810.7820.7810.780
Gradient Boost (134)0.8740.8730.8730.8740.928

Classification report of the various machine learning models for 10-fold.

4.2.4 Explainable AI models

The Random Forest model is selected to explain the LIME and SHAPELY models of the XAI. The LIME model explains the local surrogates and estimates which features are positive (increase) and which are negative toward the prediction of the target class. This model is used in a local surrogate for a particular dataset instance. This application also determines the feature weights and prediction score for each classifier in accordance with a specific instance.

SHAPELY uses various models based on the explainer suggested by Random Forest. It provides the testpatch, which distributes features in the global surrogates. Then, SHAPELY uses plots like summary plot, which provides the order of the features that determine the magnitude of the output. It also provides the dependency plot, which explains the dependency between the two variables of interest in global surrogacy. The decision plot of SHAPELY provides the decision on a particular instance and explains the rationale behind the classification with the feature impact analysis.

The first model discussed for explainability is the partial dependency plot (PDP). This plot shows the relationship between the two contributing features through linear relationship estimation through LASSO. The correlation between the two attributes is represented by the PDP. The plot between the MaxHR with the target feature Heart Disease is presented by the PDP plot in

Figure 8

. The LIME model predicts the behavior of an instance in the local surrogacy and explains the relationship between the target attribute and the rest of the features in the dataset. This also estimates the attribute weights, which features provide a positive relationship to the target prediction, and which features provide a negative response. According to an instance depicted in

Figure 9

, class 0, which is no disease, has a 2% probability, and class 1, which is the Heart Disease, has a 98% probability of occurrence. This notebook model explains the list of the features that influence the target attribute.

Figure 10

shows the Pyplot, which describes the features that have a positive relationship towards the target, such as 1_slope, Chest Pain Type, Age, Cholesterol, Blood Sugar, Exercise Angina, Sex, and MaxHR. The features with a negative relationship to the target, like Old Peak and Blood Sugar, are also explained. Using linear relationships, LIME thus explains the relationship between the target attribute and the rest of the attributes in a particular row instance. This also estimates the feature weight, nature, and significance of that particular local surrogacy. The SHAPELY explainer provides local and global surrogate explanations for the local instance and the complete dataset, respectively. It uses various plots to describe each feature’s significance in determining the target’s magnitude. The plots that are depicted in this work include

  • Force plot

  • Test patch

  • Dependency plot

  • Summary plot

  • Decision plot

The force plot explains an instance in the local surrogacy and tells how the feature values take a range between minimum and maximum, with the perception of a corresponding instance. It shows how the features contribute to the model prediction for a specific observation, as shown in

Figure 11

. The prediction score for this model is 0.98. The red-colored features increase the prediction score, and the blue-colored features decrease the prediction score. The features closer to this dividing region have the highest impact on the model prediction for that particular instance. In this instance, the parameter Cholesterol is for increasing the prediction score and 1_slope for decreasing the prediction. The classic test patch provides the overall distribution of features and shows how they can help predict the target. This global surrogate model explains the entire dataset regarding what features contribute to the prediction of heart failure through a double-colored area. The red color shows chances for Heart Failure, and the blue shows normal output. The classy test patch is described in

Figure 12

. In this plot, the features closer to the dividing boundary are also highly important in predicting the model. The summary plot lists various features in the dataset and sorts them based on the order of significance in determining the magnitude of the output. Cholesterol, Maximum Heart Rate, Blood Pressure, Age, and Chest Pain Type have the order of significance in determining the target value, respectively. The features and their corresponding weight importance are shown in

Figure 13

.

Figure 14

depicts the summary plot with feature concentration. The target value, Heart Failure, is distributed from 0 to 1. Various features like 1_slope, Chest Pain Type, Exercise Angina, Old Peak, and Cholesterol are plotted as per the order of significance in determining the output magnitude. The red-colored region shows a high impact, and the blue-colored region shows a low impact in predicting the target attribute. The SHAPELY decision plot is illustrated in

Figure 15

. This is a global surrogate model, where the dependency between the target class and the cholesterol is plotted in the graph in

Figure 15

. PDP also looks similar to the dependency plot of SHAPELY, but SHAPELY provides granular outputs that can be increased or minimized. The second point is that PDP is only a plot, but a dependency plot is a variable-like result. Taking the average value per variable is like plotting variable importance against the SHAP value, which will look like a PDP graph.

Figure 8

Figure 9

Figure 10

Figure 11

Figure 12

Figure 13

Figure 14

Figure 15

5 Discussion

This section deals with the comparative analysis of various machine learning algorithms that are used in this work. This work also deals with how the features contribute to the results in the SHAPELY explainer. The comparative analysis of the various machine learning algorithms is presented in Figure 16. The ratio of rightly predicted data to the total observations determines the accuracy of the model. The ratio of the rightly predicted positive data to the total analyzed positives fixes the precision. The ratio of the rightly predicted positive data to all actual positives is a recall metric. F1-score defines the harmonic mean of both precision and recall. The Random Forest model, which has a higher accuracy of 0.955 and F1-score of 0.955, was selected for explanation by XAI applications. The second-best values for accuracy and F1-score are recorded in the Gradient Boost model with values of 0.935 and 0.935 with a precision of 0.997. Logistic Regression and SVM have accuracy values of 0.875 and 0.875. All these models only have marginal differences in the values of parameters between them. The Decision Tree model recorded a highest accuracy of 0.961, but the AUC was the highest for random forest, which is 0.994. Thus, this model is selected for XAI implementation.

Figure 16

The 10-fold validation is also presented in Figure 17. These results show reduced accuracy levels with the lack of standard preprocessing techniques. Despite the reduced accuracy levels, Gradient Boosting and Random Forest algorithms perform much better than the rest of the models. The SHAPELY decision plots are presented in Figures 18, 19. These decision plots are extremely important in determining why an instance is classified as normal or abnormal (Heart Failure). In this local instance, the values of 1_slope, ECG Peak, and Exercise Angina are high. The value of cholesterol is also high, and the Chest Pain Type is Recorded as Type 3. All these feature values correspond to the heart disease classification into 1, which means a risk indication of Heart Failure. In the case of Figure 19, all the feature values are normal, and the instance is classified into the normal category. Thus, the decision plot of SHAPELY values provides a detailed explanation regarding how an instance is classified on the basis of various values of the features available.

Figure 17

Figure 18

Figure 19

5.1 Challenges

This work has the following challenges (not limited to), which are required to be addressed in the future. The sensors may go out of order and hence can provide false alarms to the cloud and database. The electronic faults may induce false alarms regarding heart failure. The medical data are subjected to be private. Explaining may compromise the privacy and integrity of the individual medical data. Medical data stored in the cloud are vulnerable to attacks if no security mechanisms are provided. If the medical record is stored in a blockchain model, it is extremely difficult to access and explain the same with the XAI model. The reliability of the explanation and privacy need to be enhanced by Federated Learning. Training and demonstration are required for medical practitioners to handle data from wearable sensors and the cloud.

5.2 Contributions of the paper

The essential contributions of this paper helps identify the complete purpose of this research. This paper provides a complete illustration of all the sections of IoMT-enabled XAI infrastructure. It also discusses various IoMT applications and case studies related to heart failure in detail. This paper works with a dataset with all the vital parameters required for heart failure prediction. It provides solutions for the explanation of heart failure through local and global surrogates with the explanations of LIME and SHAPELY. This study discusses various state-of-the-art IoMT sensors with practical applicability in medical applications with a discussion of advantages and disadvantages.

5.3 Future work

Improvements can be made to this study by applying many advanced techniques. The application of 6G may improve the connectivity and network-related issues associated with wearable sensors. Application of Federated Learning would improve the privacy, reliability, and safety of medical data. Meta-verse applications can enhance IoMT sensor support and provide real-time solutions to heart problems. Industry 5.0 can enhance the quality of service of the proposed system with a human-centric man–machine interface. Web 3.0 standards can provide better semantics, security, and reliability in cloud service.

6 Conclusion

Early detection of heart failure is the most desirable and need-of-the-hour application, as the number of cardiac arrest cases increases day by day. A healthy life cycle, clean habits, and a peaceful life are the real medicines to overcome heart disease. Clinical efforts are merely supplementary but not primary in nature in addressing the issues related to heart failure. The IoMT integrated Heart Failure prediction model discussed in this study is extremely useful in this stressful modern-day life. The IoMT sensors can control and monitor most of the parameters relevant to heart failure at the primary level. XAI provides excellent support to this system by indicating what body parameters influence the heart failure condition through various models that show the significance of the features for the prediction of the target. The probability of prediction of the Random Forest model is used by LIME, the local explainer, and SHAPELY, the global explainer, for explaining models related to heart failure prediction. This model is a whistleblower to many such systems developed to make human life longer, better, and safer.

Statements

Data availability statement

The original contributions presented in the study are included in the article/Supplementary Material, further inquiries can be directed to the corresponding author.

Ethics statement

Ethical review and approval was not required for the study on human participants in accordance with the local legislation and institutional requirements. All datasets used in this research are publicly available and have been utilized in compliance with their respective terms of use and ethical guidelines.

Author contributions

NK: Conceptualization, Methodology, Software, Investigation, Writing – original draft. GE: Investigation, Data curation, Methodology, Software, Conceptualization, Writing – original draft. SS: Writing – review & editing, Supervision, Project administration, Validation, Visualization. RKD: Project administration, Validation, Supervision, Formal analysis, Writing – original draft, Visualization. DP: Visualization, Formal analysis, Validation, Project administration, Supervision, Writing – review & editing. NS: Project administration, Writing – review & editing, Validation, Supervision, Formal analysis, Visualization.

Funding

The author(s) declare that no financial support was received for the research and/or publication of this article.

Conflict of interest

The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.

Generative AI statement

The author(s) declare that no Generative AI was used in the creation of this manuscript.

Any alternative text (alt text) provided alongside figures in this article has been generated by Frontiers with the support of artificial intelligence and reasonable efforts have been made to ensure accuracy, including review by the authors wherever possible. If you identify any issues, please contact us.

Publisher’s note

All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.

References

  • 1.

    TsaoCWAdayAWAlmarzooqZIAndersonCAAroraPAveryCLet alHeart disease and stroke statistics 2023 update: a report from the American Heart Association. Circulation. (2023) 147:e93621. 10.1161/CIR.0000000000001123

  • 2.

    NabelEG. Cardiovascular disease. New Engl J Med. (2003) 349:6072. 10.1056/NEJMra035098

  • 3.

    PharmDBergerK. Data from: Heart Disease Statistics 2025 (2025). Available online at:https://www.singlecare.com/blog/news/heart-disease-statistics/ (Accessed April 20, 2023).

  • 4.

    GazianoTReddyKSPaccaudFHortonSChaturvediV. Cardiovascular disease. In: Jamison DT, Breman JG, Measham AR, Alleyne G, Claeson M, Evans DB, et al., editors. Disease Control Priorities in Developing Countries. 2nd ed.Washington DC: The International Bank for Reconstruction and Development/The World Bank (2006). p. 64562. Available online at: https://www.ncbi.nlm.nih.gov/pubmed/21250342

  • 5.

    KhatibzadehSFarzadfarFOliverJEzzatiMMoranA. Worldwide risk factors for heart failure: a systematic review and pooled analysis. Int J Cardiol. (2013) 168:118694. 10.1016/j.ijcard.2012.11.065

  • 6.

    UeshimaHSekikawaAMiuraKTurinTCTakashimaNKitaYet alCardiovascular disease and risk factors in Asia: a selected review. Circulation. (2008) 118:27029. 10.1161/CIRCULATIONAHA.108.790048

  • 7.

    World Health Organization. Global Action Plan for the Prevention and Control of NCDs 2013-2020. Geneva: World Health Organization (2014). Available online at: https://www.who.int/publications/i/item/9789241506236

  • 8.

    GuoTFanYChenMWuXZhangLHeTet alCardiovascular implications of fatal outcomes of patients with coronavirus disease 2019 (COVID-19). JAMA Cardiol. (2020) 5:8118. 10.1001/jamacardio.2020.1017

  • 9.

    ShiSQinMShenBCaiYLiuTYangFet alAssociation of cardiac injury with mortality in hospitalized patients with COVID-19 in Wuhan, China. JAMA Cardiol. (2020) 5:80210. 10.1001/jamacardio.2020.0950

  • 10.

    GuptaAMadhavanMVSehgalKNairNMahajanSSehrawatTSet alExtrapulmonary manifestations of COVID-19. Nat Med. (2020) 26:101732. 10.1038/s41591-020-0968-3

  • 11.

    XieYBoweBMaddukuriGAl-AlyZ. Comparative evaluation of clinical manifestations and risk of death in patients admitted to hospital with COVID-19 and seasonal influenza: cohort study. BMJ. (2020) 371:m4677. 10.1136/bmj.m4677

  • 12.

    BoehmerTKKompaniyetsLLaveryAMHsuJKoJYYusufHet alAssociation between COVID-19 and myocarditis using hospital-based administrative data—United States, March 2020–January 2021. Morb Mortal Wkly Rep. (2021) 70:122832. 10.15585/mmwr.mm7035e5

  • 13.

    PatoneMMeiXWHandunnetthiLDixonSZaccardiFShankar-HariMet alRisks of myocarditis, pericarditis, and cardiac arrhythmias associated with COVID-19 vaccination or SARS-CoV-2 infection. Nat Med. (2022) 28:41022. 10.1038/s41591-021-01630-0

  • 14.

    BlockJPBoehmerTKForrestCBCartonTWLeeGMAjaniUAet alCardiac complications after SARS-CoV-2 infection and mRNA COVID-19 vaccination—PCORnet, United States, January 2021–January 2022. Morb Mortal Wkly Rep. (2022) 71:517. 10.15585/mmwr.mm7114e1

  • 15.

    KimY-EHuhKParkY-JPeckKRJungJ. Association between vaccination and acute myocardial infarction and ischemic stroke after COVID-19 infection. JAMA. (2022) 328:8879. 10.1001/jama.2022.12992

  • 16.

    WitbergGBardaNHossSRichterIWiessmanMAvivYet alMyocarditis after COVID-19 vaccination in a large health care organization. New Engl J Med. (2021). 385(23):21329. 10.1056/nejmoa2110737

  • 17.

    KarlstadØHoviPHusbyAHärkänenTSelmerRMPihlströmNet alSARS-CoV-2 vaccination and myocarditis in a Nordic cohort study of 23 million residents. JAMA Cardiol. (2022) 7:60012. 10.1001/jamacardio.2022.0583

  • 18.

    LingRRRamanathanKTanFLTaiBCSomaniJFisherDet alMyopericarditis following COVID-19 vaccination and non-COVID-19 vaccination: a systematic review and meta-analysis. Lancet Respir Med. (2022) 10(7):67988. 10.1016/s2213-2600(22)00059-5

  • 19.

    World Health Organization. Data from: COVID-19 vaccine heart disease: risks, safety, and recommendations. Available online at: https://www.who.int/emergencies/diseases/novel-coronavirus-2019/covid-19-vaccines/advice (Accessed August 15, 2025).

  • 20.

    MyersTRMarquezPLGeeJMHauseAMPanagiotakopoulosLZhangBet alThe V-safe after vaccination health checker: active vaccine safety monitoring during CDC’s COVID-19 pandemic response. Vaccine. (2023) 41(7):13108. 10.1016/j.vaccine.2022.12.031

  • 21.

    OtoomAFAbdallahEEKilaniYKefayeAAshourM. Effective diagnosis and monitoring of heart disease. Int J Software Eng Appl. (2015) 9:14356.

  • 22.

    PustokhinaIVPustokhinDAGuptaDKhannaAShankarKNguyenGN. An effective training scheme for deep neural network in edge computing enabled internet of medical things (IoMT) systems. IEEE Access. (2020) 8:10711223. 10.1109/ACCESS.2020.3000322

  • 23.

    ManickamPMariappanSAMurugesanSMHansdaSKaushikAShindeRet alArtificial intelligence (AI) and internet of medical things (IoMT) assisted biomedical systems for intelligent healthcare. Biosensors. (2022) 12:562. 10.3390/bios12080562

  • 24.

    ManogaranGChilamkurtiNHsuC-H. Emerging trends, issues, and challenges in internet of medical things and wireless networks. Pers Ubiquitous Comput. (2018) 22:87982. 10.1007/s00779-018-1178-6

  • 25.

    Moreno-SanchezPA. Improvement of a prediction model for heart failure survival through explainable artificial intelligence. arXiv [Preprint]. arXiv:2108.10717 (2021).

  • 26.

    OnianiSMarquesGBarnoviSPiresIMBhoiAK. Artificial intelligence for Internet of things and enhanced medical systems. In: Bhoi A, Mallick PK, Liu C-M, Balas VE, editors. Studies in Computational Intelligence. Singapore: Springer (2020) p. 4359. 10.1007/978-981-15-5495-7_3

  • 27.

    Pratap SinghRJavaidMHaleemAVaishyaRAliS. Internet of medical things (IoMT) for orthopaedic in COVID-19 pandemic: roles, challenges, and applications. J Clin Orthop Trauma. (2020) 11(4):7137. 10.1016/j.jcot.2020.05.011

  • 28.

    WrazenWGontarskaKGrzelkaFPolzeA. Explainable AI for medical event prediction for heart failure patients. In: International Conference on Artificial Intelligence in Medicine. Slovania: Springer (2023). p. 97–107.

  • 29.

    MathurPSrivastavaSXuXMehtaJL. Artificial intelligence, machine learning, and cardiovascular disease. Clin Med Insights Cardiol. (2020) 14:117954682092740. 10.1177/1179546820927404

  • 30.

    ChangVBhavaniVRXuAQHossainM. An artificial intelligence model for heart disease detection using machine learning algorithms. Healthc Anal. (2022) 2:100016. 10.1016/j.health.2022.100016

  • 31.

    LohHWOoiCPSeoniSBaruaPDMolinariFAcharyaUR. Application of explainable artificial intelligence for healthcare: a systematic review of the last decade (2011–2022). Comput Methods Programs Biomed. (2022) 226:107161. 10.1016/j.cmpb.2022.107161

  • 32.

    SrinivasuPNSandhyaNJhaveriRHRautR. From blackbox to Explainable AI in healthcare: existing tools and case studies. Mobile Inf Syst. (2022) 2022:120. 10.1155/2022/8167821

  • 33.

    GerlingsJJensenMSSholloA. Explainable AI, but explainable to whom? An exploratory case study of XAI in healthcare. In: Lim CP, Chen YW, Vaidya A, Mahorkar C, Jain LC, editors. Handbook of Artificial Intelligence in Healthcare: Vol 2: Practicalities and Prospects. Denmark: Springer (2022). p. 169–98.

  • 34.

    RazaATranKPKoehlLLiS. Designing ECG monitoring healthcare system with federated transfer learning and explainable AI. Knowl Based Syst. (2022) 236:107763. 10.1016/j.knosys.2021.107763

  • 35.

    KumarRGargSKaurRJoharMGMSinghSMenonSVet alA comprehensive review of machine learning for heart disease prediction: challenges, trends, ethical considerations, and future directions. Front Artif Intell. (2025) 8:1583459. 10.3389/frai.2025.1583459

  • 36.

    RajkumarGGayathri DeviTSrinivasanA. Heart disease prediction using IoT based framework and improved deep learning approach: medical application. Med Eng Phys. (2023) 111:103937. 10.1016/j.medengphy.2022.103937

  • 37.

    HammadiJFLatifABACobZBC. An improved framework for cardiovascular disease prediction using hybrid ensemble learning soft-voting model. Int J Comput Exp Sci Eng. (2025) 11(3):386270. 10.22399/ijcesen.2348

  • 38.

    RokoniSIslamMAArafuzzamanMBasharSHasanMRIslamMT. Heart disease prediction: leveraging CNNs and XAI modeling for interpretability. J Indep Stud Res Comput. (2025) 23:5257. 10.31645/JISRC.25.23.1.6

  • 39.

    WangXSongY. Edge-assisted IoMT-based smart-home monitoring system for the elderly with chronic diseases. IEEE Sens Lett. (2023) 7:14. 10.1109/LSENS.2023.3240670

  • 40.

    MartinekRTiwariPVidyarthiAAlkhayyatAWangW. Federated-learning based privacy preservation and fraud-enabled blockchain IoMT system for healthcare. IEEE J Biomed Health Inform. (2023) 27(2):66472. 10.1109/JBHI.2022.3165945

  • 41.

    KumarMVermaSKumarAIjazMFRawatDBetal. ANAF-IoMT: a novel architectural framework for IoMT-enabled smart healthcare system by enhancing security based on RECC-VC. IEEE Trans Ind Inform. (2022) 18:893643. 10.1109/TII.2022.3181614

  • 42.

    KishorAChakrabortyC. Artificial intelligence and internet of things based healthcare 4.0 monitoring system. Wirel Pers Commun. (2022) 127:161531. 10.1007/s11277-021-08708-5

  • 43.

    ShafiqMDuCJamalNAbroJHKamalTAfsarSet alSmart e-health system for heart disease detection using artificial intelligence and internet of things integrated next-generation sensor networks. J Sens. (2023) 2023(1):17. 10.1155/2023/6383099

  • 44.

    KumarPMGandhiUD. A novel three-tier internet of things architecture with machine learning algorithm for early detection of heart diseases. Comput Electr Eng. (2018) 65:22235. 10.1016/j.compeleceng.2017.09.001

  • 45.

    PanjaSChattopadhyayAKNagASinghJP. Fuzzy-logic-based IoMT framework for COVID19 patient monitoring. Comput Ind Eng. (2023) 176:108941. 10.1016/j.cie.2022.108941

  • 46.

    JainSNehraMKumarRDilbaghiNHuTKumarSet alInternet of medical things (IoMT)-integrated biosensors for point-of-care testing of infectious diseases. Biosens Bioelectron. (2021) 179:113074. 10.1016/j.bios.2021.113074

  • 47.

    RezaeeKYangXKhosraviMRZhangRLinWJeonG. Fusion-based learning for stress recognition in smart home: an IoMT framework. Build Environ. (2022) 216:108988. 10.1016/j.buildenv.2022.108988

  • 48.

    KrupaAJDDhanalakshmiSLaiKWTanYWuX. An IoMT enabled deep learning framework for automatic detection of fetal QRS: a solution to remote prenatal care. J King Saud Univ Comput Inf Sci. (2022) 34:720011. 10.1016/j.jksuci.2022.07.002

  • 49.

    LuYQiYFuX. A framework for intelligent analysis of digital cardiotocographic signals from IoMT-based foetal monitoring. Future Gen Comput Syst. (2019) 101:113041. 10.1016/j.future.2019.07.052

  • 50.

    RahmaniAMGiaTNNegashBAnzanpourAAzimiIJiangMet alExploiting smart e-Health gateways at the edge of healthcare Internet-of-Things: a fog computing approach. Future Gen Comput Syst. (2018) 78:64158. 10.1016/j.future.2017.02.014

  • 51.

    NandySAdhikariMChakrabortySAlkhayyatAKumarN. IBoNN: Intelligent Agent-based Internet of Medical Things framework for detecting brain response from electroencephalography signal using bag-of-neural network. Future Gen Comput Syst. (2022) 130:24152. 10.1016/j.future.2021.12.019

  • 52.

    YadavAVermaDKumarAKumarPSolankiP. The perspectives of biomarker-based electrochemical immunosensors, artificial intelligence and the Internet of Medical Things toward COVID-19 diagnosis and management. Mater Today Chem. (2021) 20:100443. 10.1016/j.mtchem.2021.100443

  • 53.

    VermaDSinghKRYadavAKNayakVSinghJSolankiPRet alInternet of things (IoT) in nano-integrated wearable biosensor devices for healthcare applications. Biosens Bioelectron X. (2022) 11:100153. 10.1016/j.biosx.2022.100153

  • 54.

    FouadHHassaneinASSolimanAMAl-FeelH. Internet of medical things (IoMT) assisted vertebral tumor prediction using heuristic hock transformation based Gautschi model–a numerical approach. IEEE Access. (2020) 8:17299309. 10.1109/ACCESS.2020.2966272

  • 55.

    RibeiroMTSinghSGuestrinC. Why should I trust you? In: Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining. San Francisco: ACM (2016).

  • 56.

    LundbergSMLeeS-I. A unified approach to interpreting model predictions. arXiv [Preprint]. arXiv:1705.07874 (2017). 10.48550/ARXIV.1705.07874

  • 57.

    MuddamsettySMJahromiMNMoeslundTB. Expert level evaluations for explainable AI (XAI) methods in the medical domain. In: Pattern Recognition. ICPR International Workshops and Challenges: Virtual Event, January 10–15, 2021, Proceedings, Part III. Berlin, Heidelberg: Springer (2021). p. 35–46.

  • 58.

    TjoaEGuanC. A survey on explainable artificial intelligence (XAI): Toward medical XAI. IEEE Trans Neural Networks Learn Syst. (2020) 32:4793813. 10.1109/TNNLS.2020.3027314

  • 59.

    OnanA. Hierarchical graph-based text classification framework with contextual node embedding and BERT-based dynamic fusion. J King Saud Univ Comput Inf Sci. (2023) 35:101610. 10.1016/j.jksuci.2023.101610

  • 60.

    OnanA. GTR-GA: harnessing the power of graph-based neural networks and genetic algorithms for text augmentation. Expert Syst Appl. (2023) 232:120908. 10.1016/j.eswa.2023.120908

  • 61.

    OnanA. SRL-ACO: a text augmentation framework based on semantic role labeling and ant colony optimization. J King Saud Univ Comput Inf Sci. (2023) 35:101611. 10.1016/j.jksuci.2023.101611

  • 62.

    OnanA. Bidirectional convolutional recurrent neural network architecture with group-wise enhancement mechanism for text sentiment classification. J King Saud Univ Comput Inf Sci. (2022) 34:2098117. 10.1016/j.jksuci.2022.02.025

  • 63.

    OnanA. Two-stage topic extraction model for bibliometric data analysis based on word embeddings and clustering. IEEE Access. (2019) 7:14561433. 10.1109/ACCESS.2019.2945911

  • 64.

    OnanAKorukoğluSBulutH. Ensemble of keyword extraction methods and classifiers in text classification. Expert Syst Appl. (2016) 57:23247. 10.1016/j.eswa.2016.03.045

  • 65.

    OnanA. Consensus clustering-based undersampling approach to imbalanced learning. Sci Program. (2019) 2019:114. 10.1155/2019/5901087

  • 66.

    OnanAKorukoğluS. A feature selection model based on genetic rank aggregation for text sentiment classification. J Inf Sci. (2017) 43:2538. 10.1177/0165551515613226

  • 67.

    OnanAKorukoğluSBulutH. A hybrid ensemble pruning approach based on consensus clustering and multi-objective evolutionary algorithm for sentiment classification. Inf Process Manage. (2017) 53:81433. 10.1016/j.ipm.2017.02.008

  • 68.

    OnanA. Sentiment analysis on product reviews based on weighted word embeddings and deep neural networks. Concurr Comput Pract Exp. (2021) 33:e5909. 10.1002/cpe.5909

  • 69.

    OnanA. Sentiment analysis on massive open online course evaluations: a text mining and deep learning approach. Comput Appl Eng Educ. (2021) 29:57289. 10.1002/cae.22253

  • 70.

    OnanA. Mining opinions from instructor evaluation reviews: a deep learning approach. Comput Appl Eng Educ. (2020) 28:11738. 10.1002/cae.22179

  • 71.

    OnanA. An ensemble scheme based on language function analysis and feature engineering for text genre classification. J Inf Sci. (2018) 44:2847. 10.1177/0165551516677911

  • 72.

    OnanAToçoğluMA. A term weighted neural language model and stacked bidirectional LSTM based framework for sarcasm identification. IEEE Access. (2021) 9:770122. 10.1109/ACCESS.2021.3049734

  • 73.

    VakhariaVShahMNairPBoradeHSahlotPWankhedeV. Estimation of lithium-ion battery discharge capacity by integrating optimized explainable-AI and stacked LSTM model. Batteries. (2023) 9:125. 10.3390/batteries9020125

  • 74.

    AliYAAwwadEMAl-RazganMMaaroufA. Hyperparameter search for machine learning algorithms for optimizing the computational complexity. Processes. (2023) 11:349. 10.3390/pr11020349

  • 75.

    SutharVVakhariaVPatelVKShahM. Detection of compound faults in ball bearings using multiscale-SinGAN, heat transfer search optimization, and extreme learning machine. Machines. (2023) 11:29. 10.3390/machines11010029

  • 76.

    VakhariaVGuptaVKKankarPK. A comparison of feature ranking techniques for fault diagnosis of ball bearing. Soft Comput. (2016) 20:160119. 10.1007/s00500-015-1608-6

  • 77.

    HashemMVellappallySFouadHLuqmanMYoussefAE. Predicting neurological disorders linked to oral cavity manifestations using an IoMT-based optimized neural networks. IEEE Access. (2020) 8:19072233. 10.1109/ACCESS.2020.3027632

  • 78.

    ZhuTKuangLDanielsJHerreroPLiKGeorgiouP. IoMT-enabled real-time blood glucose prediction with deep learning and edge computing. IEEE Internet Things J. (2023) 10(5):370619. 10.1109/JIOT.2022.3143375

  • 79.

    KhanTAFatimaAShahzadTAtta-Ur-Rahman, AlissaKGhazalTMet alSecure IoMT for disease prediction empowered with transfer learning in healthcare 5.0, the concept and case study. IEEE Access. (2023) 11:3941830. 10.1109/access.2023.3266156

  • 80.

    NandySAdhikariMKhanMAMenonVGVermaS. An intrusion detection mechanism for secured IoMT framework based on swarm-neural network. IEEE J Biomed Health Inf. (2021) 26:196976. 10.1109/JBHI.2021.3101686

  • 81.

    LakhanAMohammedMANedomaJMartinekRTiwariPVidyarthiAet alFederated-learning based privacy preservation and fraud-enabled blockchain IoMT system for healthcare. IEEE J Biomed Health Inform. (2023) 27(2):66472. 10.1109/jbhi.2022.3165945

  • 82.

    ZhangTSodhroAHLuoZZahidNNawazMWPirbhulalSet alA joint deep learning and Internet of Medical Things driven framework for elderly patients. IEEE Access. (2020) 8:7582232. 10.1109/ACCESS.2020.2989143

  • 83.

    KhanMAAlgarniF. A healthcare monitoring system for the diagnosis of heart disease in the IoMT cloud environment using MSSO-ANFIS. IEEE Access. (2020) 8:12225969. 10.1109/ACCESS.2020.3006424

  • 84.

    GuleriaPNaga SrinivasuPAhmedSAlmusallamNAlarfajFK. XAI framework for cardiovascular disease prediction using classification techniques. Electronics. (2022) 11:4086. 10.3390/electronics11244086

  • 85.

    JuholaMJoutsijokiHPenttinenKAalto-SetäläK. Detection of genetic cardiac diseases by Ca2+ transient profiles using machine learning methods. Sci Rep. (2018) 8:9355. 10.1038/s41598-018-27695-5

  • 86.

    LongNCMeesadPUngerH. A highly accurate firefly based algorithm for heart disease prediction. Expert Syst Appl. (2015) 42:822131. 10.1016/j.eswa.2015.06.024

  • 87.

    PalaniappanSAwangR. Intelligent heart disease prediction system using data mining techniques. In: 2008 IEEE/ACS International Conference on Computer Systems and Applications. Qatar: IEEE (2008). p. 108–15.

  • 88.

    JinBCheCLiuZZhangSYinXWeiX. Predicting the risk of heart failure with EHR sequential data modeling. IEEE Access. (2018) 6:925661. 10.1109/ACCESS.2017.2789324

  • 89.

    PuyalnithiTViswanathamVM. Preliminary cardiac disease risk prediction based on medical and behavioural data set using supervised machine learning techniques. Indian J Sci Technol. (2016) 9:15. 10.17485/ijst/2016/v9i31/96740

  • 90.

    ForssenHPatelRFitzpatrickNHingoraniATimmisAHemingwayHet alEvaluation of machine learning methods to predict coronary artery disease using metabolomic data. Stud Health Technol Inform. (2017) 235:111–5.

  • 91.

    TangZ-HLiuJZengFLiZYuXZhouL. Comparison of prediction model for cardiovascular autonomic dysfunction using artificial neural network and logistic regression analysis. PLoS One. (2013) 8:e70571. 10.1371/journal.pone.0070571

  • 92.

    ToshniwalDGoelBSharmaH. Multistage classification for cardiovascular disease risk prediction. In: Big Data Analytics: 4th International Conference, BDA 2015, Hyderabad, India, December 15–18, 2015, Proceedings 4. Hyderabad: Springer (2015). p. 258–66.

  • 93.

    YangFLuoJHouQXieNNieZHuangWet alPrediction of cardiac death after adenosine myocardial perfusion SPECT based on machine learning. J Nucl Cardiol. (2019) 26:63341. 10.1007/s12350-017-1014-9

  • 94.

    MustaqeemAAnwarSMMajidMKhanAR. Wrapper method for feature selection to classify cardiac arrhythmia. In: 2017 39th Annual International Conference of the IEEE Engineering in Medicine and Biology Society (EMBC). Jeju Island: IEEE (2017). p. 3656–9.

  • 95.

    MansoorHElgendyIYSegalRBavryAABianJ. Risk prediction model for in-hospital mortality in women with ST-elevation myocardial infarction: a machine learning approach. Heart Lung. (2017) 46:40511. 10.1016/j.hrtlng.2017.09.003

  • 96.

    KimJLeeJLeeY. Data-mining-based coronary heart disease risk prediction model using fuzzy logic and decision tree. Healthc Inform Res. (2015) 21:16774. 10.4258/hir.2015.21.3.167

  • 97.

    TaslimitehraniVDongGPereiraNLPanahiazarMPathakJ. Developing EHR-driven heart failure risk prediction models using CPXR (log) with the probabilistic loss function. J Biomed Inform. (2016) 60:2609. 10.1016/j.jbi.2016.01.009

  • 98.

    AnbarasiMAnupriyaEIyengarN. Enhanced prediction of heart disease with feature subset selection using genetic algorithm. Int J Eng Sci Technol. (2010) 2(10):53706.

  • 99.

    BhatlaNJyotiK. An analysis of heart disease prediction using different data mining techniques. Int J Eng. (2012) 1(8):14.

  • 100.

    ThenmozhiKDeepikaP. Heart disease prediction using classification with different decision tree techniques. Int J Eng Res Gen Sci. (2014) 2(6):611.

  • 101.

    TamilarasiRPorkodiDR. A study and analysis of disease prediction techniques in data mining for healthcare. Int J Emerg Res Manag Technol. (2015) 1:22789359.

  • 102.

    MarikaniTShyamalaK. Prediction of heart disease using supervised learning algorithms. Int J Comput Appl. (2017) 165(5):414. 10.5120/ijca2017913868

  • 103.

    LuPGuoSZhangHLiQWangYWangYet alResearch on improved depth belief network-based prediction of cardiovascular diseases. J Healthc Eng. (2018) 2018:19. 10.1155/2018/8954878

  • 104.

    RaoAEskandar-AfshariFWeinerYBillmanEMcMillinASellaNet alClinical study of continuous non-invasive blood pressure monitoring in neonates. Sensors. (2023) 23:3690. 10.3390/s23073690

  • 105.

    JakobVKüderleAKlugeFKluckenJEskofierBMWinklerJet alValidation of a sensor-based gait analysis system with a gold-standard motion capture system in patients with Parkinson’s disease. Sensors. (2021) 21:7680. 10.3390/s21227680

  • 106.

    RegidorP-AKaczmarczykMSchiweckEGoeckenjan-FestagMAlexanderH. Identification and prediction of the fertile window with a new web-based medical device using a vaginal biosensor for measuring the circadian and circamensual core body temperature. Gynecol Endocrinol. (2018) 34:25660. 10.1080/09513590.2017.1390737

  • 107.

    Morgado AreiaCSantosMVollamSPimentelMYoungLRomanCet alA chest patch for continuous vital sign monitoring: clinical validation study during movement and controlled hypoxia. J Med Internet Res. (2021) 23:e27547. 10.2196/27547

  • 108.

    MaqboolSParkmanHPFriedenbergFK. Wireless capsule motility: comparison of the SmartPill® GI monitoring system with scintigraphy for measuring whole gut transit. Dig Dis Sci. (2009) 54:216774. 10.1007/s10620-009-0899-9

  • 109.

    ShangTZhangJYThomasAArnoldMAVetterBNHeinemannLet alProducts for monitoring glucose levels in the human body with noninvasive optical, noninvasive fluid sampling, or minimally invasive technologies. J Diabetes Sci Technol. (2022) 16:168214. 10.1177/19322968211007212

  • 110.

    Perez-MarcosDChevalleyOSchmidlinTGaripelliGSerinoAVuadensPet alIncreasing upper limb training intensity in chronic stroke using embodied virtual reality: a pilot study. J Neuroeng Rehabil. (2017) 14:114. 10.1186/s12984-017-0328-9

  • 111.

    AmmannKRAhamedTSweedoALGhaffariRWeinerYESlepianRCet alHuman motion component and envelope characterization via wireless wearable sensors. BMC Biomed Eng. (2020) 2:115. 10.1186/s42490-020-0038-4

  • 112.

    BrezulianuABurlacuAPopaIVArifMGemanO. “Not by our feeling, but by other’s seeing”: sentiment analysis technique in cardiology—an exploratory review. Front Public Health. (2022) 10:880207. 10.3389/fpubh.2022.880207

  • 113.

    SundholmMChengJZhouBSethiALukowiczP. Smart-mat: recognizing and counting gym exercises with low-cost resistive pressure sensing matrix. In: Proceedings of the 2014 ACM International Joint Conference on Pervasive and Ubiquitous Computing. Seattle: Pervasive and Mobile Computing (2014). p. 373–82.

  • 114.

    AdlerSNMetzgerYC. Pillcam colon capsule endoscopy: recent advances and new insights. Therap Adv Gastroenterol. (2011) 4:2658. 10.1177/1756283x11401645

  • 115.

    ZhouASantacruzSRJohnsonBCAlexandrovGMoinABurghardtFLet alA wireless and artefact-free 128-channel neuromodulation device for closed-loop stimulation and recording in non-human primates. Nat Biomed Eng. (2019) 3:1526. 10.1038/s41551-018-0323-x

  • 116.

    ShuklaAKMehaniRSadasivamB. Abilify Mycite (aripiprazole): a critical evaluation of the novel dosage form. J Clin Psychopharmacol. (2021) 41:934. 10.1097/JCP.0000000000001334

  • 117.

    RegaliaGOnoratiFLaiMCaborniCPicardRW. Multimodal wrist-worn devices for seizure detection and advancing research: focus on the Empatica wristbands. Epilepsy Res. (2019) 153:7982. 10.1016/j.eplepsyres.2019.02.007

  • 118.

    DavisGMPetersALBodeBWCarlsonALDumaisBVienneauTEet alSafety and efficacy of the Omnipod 5 automated insulin delivery system in adults with type 2 diabetes: from injections to hybrid closed-loop therapy. Diabetes Care. (2023) 46:74250. 10.2337/dc22-1915

  • 119.

    DreiseitlSOhno-MachadoL. Logistic regression and artificial neural network classification models: a methodology review. J Biomed Inform. (2002) 35:3529. 10.1016/S1532-0464(03)00034-0

  • 120.

    SuthaharanS. Support vector machine. In: Machine Learning Models and Algorithms for Big Data Classification. Integrated Series in Information Systems, Vol 36. Boston, MA: Springer (2016). p. 207–35. 10.1007/978-1-4899-7641-3_9

  • 121.

    CharbutyBAbdulazeezA. Classification based on decision tree algorithm for machine learning. J Appl Sci Technol Trends. (2021) 2:208. 10.38094/jastt20165

  • 122.

    NatekinAKnollA. Gradient boosting machines, a tutorial. Front Neurorobot. (2013) 7:21. 10.3389/fnbot.2013.00021

  • 123.

    LiuYWangYZhangJ. New machine learning algorithm: random forest. In: Information Computing and Applications: Third International Conference, ICICA 2012, Chengde, China, September 14–16, 2012. Proceedings 3. Chengde: Springer (2012). p. 246–52.

  • 124.

    GramegnaAGiudiciP. SHAP and LIME: an evaluation of discriminative power in credit risk. Front Artif Intell. (2021) 4:752558. 10.3389/frai.2021.752558

  • 125.

    ChromikM. Reshape: a framework for interactive explanations in XAI based on SHAP. In: Proceedings of 18th European Conference on Computer-Supported Cooperative Work. Siegen: European Society for Socially Embedded Technologies (EUSSET) (2020).

  • 126.

    Fedesoriano. Data from: Heart failure prediction dataset (2021). Available online at: https://www.kaggle.com/datasets/fedesoriano/heart-failure-prediction (Accessed April 21, 2023).

  • 127.

    RigattiSJ. Random forest. J Insur Med. (2017) 47:319. 10.17849/insm-47-01-31-39.1

  • 128.

    XuP. Review on studies of machine learning algorithms. J Phys Conf Ser. (2019) 1187:052103. 10.1088/1742-6596/1187/5/052103

  • 129.

    PurwantoADWikantikaKDeliarADarmawanS. Decision tree and random forest classification algorithms for mangrove forest mapping in Sembilang National Park, Indonesia. Remote Sens. (2022) 15:16. 10.3390/rs15010016

  • 130.

    DieberJKirraneS. Why model why? Assessing the strengths and limitations of lime. arXiv [Preprint]. arXiv:2012.00093 (2020).

  • 131.

    LaValleyMP. Logistic regression. Circulation. (2008) 117:23959. 10.1161/CIRCULATIONAHA.106.682658

  • 132.

    HuangSCaiNPachecoPPNarrandesSWangYXuW. Applications of support vector machine (SVM) learning in cancer genomics. Cancer Genom Proteom. (2018) 15(1):4151. 10.21873/cgp.20063

  • 133.

    SongY-YYingL. Decision tree methods: applications for classification and prediction. Shanghai Arch Psychiatry. (2015) 27:1305. 10.11919/j.issn.1002-0829.215044

  • 134.

    ZemelRPitassiT. A gradient-based boosting algorithm for regression problems. In: Proceedings of the 14th International Conference on Neural Information Processing Systems (NIPS'00). Cambridge, MA: MIT Press (2000). p. 675–81.

  • 135.

    RojasR. AdaBoost and the super bowl of classifiers a tutorial introduction to adaptive boosting TechRxiv. [Preprint] (2024). 10.36227/techrxiv.172107276.63524590/v1.

Summary

Keywords

XAI, LIME, SHAPELY, Random Forest, PDP, heart failure prediction, heart disease

Citation

Kailasanathan N, Ezhilarasan G, Selvarajan S, Dhanaraj RK, Pamucar D and Shankar N (2025) Heart disease prediction with a feature-sensitized interpretable framework for the Internet of Medical Things sensors. Front. Digit. Health 7:1612915. doi: 10.3389/fdgth.2025.1612915

Received

25 April 2025

Accepted

11 August 2025

Published

01 October 2025

Volume

7 - 2025

Edited by

Kirti Sundar Sahu, Canadian Red Cross, Canada

Reviewed by

Vito Santamato, University of Foggia, Italy

Adhe Lingga Dewi, BINUS University School of Computer Science, Indonesia

Kireet Muppavaram, Gandhi Institute of Technology and Management (GITAM), India

Updates

Copyright

*Correspondence: Dragan Pamucar

Disclaimer

All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.

Outline

Figures

Cite article

Copy to clipboard


Export citation file


Share article

Article metrics