Applied Data Science Smart Systems
Applied Data Science Smart Systems
Edited by
Jaiteg Singh
S B Goyal
Naveen Kumar
© 2025 selection and editorial matter, Jaiteg Singh, S B Goyal, Rajesh Kumar Kaushal, Naveen Kumar and
Sukhjit Singh Sehra; individual chapters, the contributors
The right of Jaiteg Singh, S B Goyal, Rajesh Kumar Kaushal, Naveen Kumar and Sukhjit Singh Sehra to be
identified as the authors of the editorial material, and of the authors for their individual chapters, has been
asserted in accordance with sections 77 and 78 of the Copyright, Designs and Patents Act 1988.
The Open Access version of this book, available at [Link], has been made available under a Creative
Commons [Attribution-Non Commercial-No Derivatives (CC-BY-NC-ND)] 4.0 license.
Any third party material in this book is not included in the OA Creative Commons license, unless indicated otherwise
in a credit line to the material. Please direct any permissions enquiries to the original rightsholder.
For permission to photocopy or use material electronically from this work, access [Link]
or contact the Copyright Clearance Center, Inc. (CCC), 222 Rosewood Drive, Danvers, MA 01923,
978-750-8400. For works that are not available on CCC please contact mpkbookspermissions@[Link]
Trademark notice: Product or corporate names may be trademarks or registered trademarks, and are used
only for identification and explanation without intent to infringe.
DOI: 10.1201/9781003471059
Preface xx
Chapter 6 Development of an analytical model of drain current for junctionless GAA MOSFET
including source/drain resistance 43
Amrita Kumari, Jhuma Saha, Ashish Saini and Amit Kumar
Chapter 9 GAI in healthcare system: Transforming research in medicine and care for patients 63
Mahesh A., Angelin Rosy M., Vinodh Kumar M., Deepika P., Sakthidevi I. and
Sathish C.
Chapter 11 Blood bank mobile application of IoT-based android studio for COVID-19 76
Basetty Mallikarjuna, Sandeep Bhatia, Neha Goel, and Bharat Bhushan Naib
Chapter 12 Selection of effective parameters for optimizing software testing effort estimation 82
Vikas Chahar and Pradeep Kumar Bhatia
Chapter 14 An overview of wireless sensor networks applications, challenges and security attacks 98
N. Sharmila Banu, [Link] and [Link]
Chapter 15 Internet of health things-enabled monitoring of vital signs in hospitals of the future 108
Amit Sundas, Sumit Badotra, Gurpreet Singh and Amit Verma
Chapter 16 Artificial intelligence-based learning techniques for accurate prediction and classification
of colorectal cancer 114
Yogesh Kumar, Shapali Bansal, Ankush Jariyal and Apeksha Koul
Chapter 17 SLODS: Real-time smart lane detection and object detection system 120
Tanuja Satish Dhope, Pranav Chippalkatti, Sulakshana Patil, Vijaya Gopalrao Rajeshwarkar
and Jyoti Ramesh Gangane
Chapter 18 Computational task off-loading using deep Q-learning in mobile edge computing 129
Tanuja Satish Dhope, Tanmay Dikshit, Unnati Gupta and Kumar Kartik
Chapter 20 Issues with existing solutions for grievance redressal systems and mitigation approach using
blockchain network 140
Harish Kumar, Rajesh Kumar Kaushal and Naveen Kumar
Chapter 21 A systematic approach to implement hyperledger fabric for remote patient monitoring 147
Shilpi Garg, Rajesh Kumar Kaushal and Naveen Kumar
Chapter 22 Developing spell check and transliteration tools for Indian regional language – Kannada 152
Chandrika Prasad, Jagadish S. Kallimani, Geetha Reddy and Dhanashekar K.
Chapter 25 Exploring recession indicators: Analyzing social network platforms and newspapers textual
datasets 180
Nikita Mandlik, Kanishk Barhanpurkar, Harshad Bhandwaldar, S. B. Goyal, Anand Singh
Rajawat and Surabhi Rane
Chapter 26 NIRF rankings’ effects on private engineering colleges for improving India’s educational
system looked at using computational approaches 187
Ankita Mitra, Subir Gupta, P. K. Dutta, S. B. Goyal, Wan Md. Afnan Bin Wan Mahmood
and Baharu Bin Kemat
Chapter 28 Drowsiness detection in drivers: A machine learning approach using hough circle
classification algorithm for eye retina images 202
J. Viji Gripsy, N. A. Sheela Selvakumari, S. Sahul Hameed and M. Jamila Begam
Chapter 29 Optimizing congestion collision using effective rate control with data aggregation algorithm
in wireless sensor network 209
K. Deepa, C. Arunpriya and M. Sasikala
Chapter 30 DDoS attack detection methods, challenges and opportunities: A survey 215
Jaspreet Kaur and Gurjit Singh Bhathal
Chapter 32 Optimization techniques for wireless body area network routing protocols: Analysis and
comparison 226
Swati Goel, Kalpna Guleria and Surya Narayan Panda
Chapter 33 Securing the boundless network: A comprehensive analysis of threats and exploits in
software defined network 236
Shruti Keshari, Sunil Kumar, Pankaj Kumar Sharma and Sarvesh Tanwar
Chapter 34 A bibliometric analyses on emerging trends in communication disorder 246
Muskan Chawla, Surya Narayan Panda and Vikas Khullar
Chapter 35 Enhancing latency performance in fog computing through intelligent resource allocation
and Cuckoo search optimization 256
Meena Rani, Kalpna Guleria and Surya Narayan Panda
Chapter 36 Pediatric thyroid ultrasound image classification using deep learning: A review 264
Jatinder Kumar, Surya Narayan Panda and Devi Dayal
Chapter 37 Hybrid security of EMI using edge-based steganography and three-layered cryptography 278
Divya Sharma and Chander Prabha
Chapter 38 Efficient lung cancer detection in CT scans through GLCM analysis and hybrid classification 291
Shazia Shamas, Surya Narayan Panda and Ishu Sharma
Chapter 39 Newton Raphson method for root convergence of higher degree polynomials using big
number libraries 298
Taniya Hasija, K. R. Ramkumar, Bhupendra Singh, Amanpreet Kaur and Sudesh Kumar
Mittal
Chapter 42 Exploring Image Segmentation Approaches for Medical Image Analysis 322
Rupali Pathak, Hemant Makwana and Neha Sharma
Chapter 44 Integrating metaverse and blockchain for transparent and secure logistics management 334
A.U. Nwosu, S.B. Goyal, Anand Singh Rajawat, Baharu Bin Kemat and Wan Md Afnan
Bin Wan Mahmood
Chapter 45 A systematic study of multiple cardiac diseases by using algorithms of machine learning 343
Prachi Pundhir and Dhowmya Bhatt
Chapter 46 Forecasting mobile prices: Harnessing the power of machine learning algorithms 348
Parveen Badoni, Rahul Kumar, Parvez Rahi, Ajay Pal Singh Yadav and Siroj Kumar Singh
Chapter 49 Simulation-based evaluating AODV routing protocol using wireless networks 378
Bhupal Arya, Dr. Jogendra Kumar, Dr. Parag Jain, Preeti Saroj, Mrinalinee Singh and Yogesh
Kumar
Chapter 56 Liver segmentation using shape prior features with Chan-Vese model 430
Veerpal and Jyoti Rani
Chapter 58 Sales analysis: Coca-Cola sales analysis using data mining techniques for predictions and
efficient growth in sales 448
Siddique Ibrahim S. P., Pothuri Naga Sai Saketh, Gamidi Sanjay, Bhimavarapu Charan
Tej Reddy, Mesa Ravi Kanth and Selva Kumar S.
Chapter 59 Statistical analysis of consumer attitudes towards virtual influencers in the metaverse 458
Sheetal Soni and Usha Yadav
Chapter 60 Quantum dynamics-aided learning for secure integration of body area networks within the
metaverse cybersecurity framework 467
Anand Singh Rajawat, S. B. Goyal, Jaiteg Singh and Celestine Iwendi
Chapter 64 Adaptive resource allocation and optimization in cloud environments: Leveraging machine
learning for efficient computing 499
Anand Singh Rajawat, S. B. Goyal, Manoj Kumar and Varun Malik
Chapter 65 Quantum deep learning on driven trust-based routing framework for IoT in the metaverse
context 509
S. B. Goyal, Anand Singh Rajawat, Jaiteg Singh and Chawki Djeddi
Chapter 66 Advancing network security paradigms integrating quantum computing models for
enhanced protections 517
Anand Singh Rajawat, S. B. Goyal, Chaman Verma and Jaiteg Singh
Chapter 67 Optimizing 5G and beyond networks: A comprehensive study of fog, grid, soft, and
scalable computing models 529
S. B. Goyal, Anand Singh Rajawat, Jaiteg Singh and Tony Jan
Chapter 68 Smart protocol design: Integrating quantum computing models for enhanced efficiency and
security 537
S. B. Goyal, Sugam Sharma, Anand Singh Rajawat and Jaiteg Singh
Chapter 69 Efficient IIoT framework for mitigating Ethereum attacks in industrial applications using
supervised learning with quantum classifiers 544
S. B. Goyal, Anand Singh Rajawat, Ritu Shandilya and Varun Malik
Chapter 70 Quantum computing in the era of IoT: Revolutionizing data processing and security in
connected devices 552
S. B. Goyal, Sardar M. N. Islam, Anand Singh Rajawat and Jaiteg Singh
Chapter 71 A federated learning approach to classify depression using audio dataset 560
Chetna Gupta and Vikas Khullar
Chapter 72 Securing IOT CCTV: Advanced video encryption algorithm for enhanced data protection 565
Kawalpreet Kaur, Amanpreet Kaur, Vidhyotma Gandhi and Bhupendra Singh
Chapter 74 Review of techniques for diagnosis of Meibomian gland dysfunction using IR images 576
Deepika Sood, Anshu Singla and Sushil Narang
Chapter 78 Artificial intelligence and machine vision-based assessment of rice seed quality 603
Ridhi Jindal and S. K. Mittal
List of Figures
The Second International Conference on Applied Data Science and Smart Systems (ADSSS-2023) was held
on 15-16 December 2023 at Chitkara University, Punjab, India. This multidisciplinary conference focussed
on innovation and progressive practices in science, technology, and management. The conference successfully
brought together researchers, academicians, and practitioners across different domains such as artificial intel-
ligence and machine learning, software engineering, automation, data science, business computing, data com-
munication, and computer networks. The presenters shared their most recent research works that are critical
to contemporary business and societal landscape and encouraged the participants to devise solutions for real-
world challenges.
ADSSS-2023 featured an extensive selection of tracks, each delving into critical facets of applied data science
and smart systems. “Machine Learning Principles, Smart Solutions, and Business Strategies” provided insights
into the synergy between ML principles, innovative solutions, and strategic business applications. The track on
“AI and Deep Learning” explored the latest advancements and applications in artificial intelligence and deep
learning technologies. Addressing contemporary challenges, “Data Science Techniques for Handling Epidemic,
Pandemic” showcased the role of data science in managing health crises. “Deep Intelligence for Interdisciplinary
Research” facilitated discussions on the integration of deep intelligence across diverse research domains. The
track focusing on “Software Engineering and Automation” explored methodologies that enhance efficiency and
automate processes in the realm of software development. “Data Communication and Computer Networks”
shed light on evolving communication technologies, while “Computing in Business and Learning” addressed the
intersection of computing technologies with business strategies and educational practices. Finally, “Engineering
Mathematics and Physics” provided a platform for exploring the application of mathematical and physical prin-
ciples in various engineering disciplines. This diverse array of tracks collectively contributed to a comprehensive
exploration of applied data science, fostering interdisciplinary collaboration and offering valuable insights for
future research and advancements in the field.
ADSSS-2023 was honored to host eminent scientists and researchers from across the globe, whose insightful
keynote addresses enhanced participants’ knowledge. The keynotes covered diverse topics such as AI-Powered
Quantum Cryptography, the Role of Large Language Models in Scientific Research, and the Application of Deep
Gaussian Processes in Radio Map Construction and Localization. Beyond knowledge dissemination, the confer-
ence served as a dynamic platform for networking and collaboration among researchers, fostering the exchange
of ideas that may shape future research endeavours. We trust that every participant found the ADSSS-2023
experience to be both enriching and productive.
Editors
Dr. Rajesh Kumar Kaushal is a highly accomplished and dedicated researcher and
educator, boasting a robust background in Computer Science and Engineering, and
accumulating an extensive 19 years of experience in academia since 2004. Holding a
Ph.D. in Computer Science and Engineering from Chitkara University, Punjab, India,
he currently serves as a professor in the Department of Computer Applications and
actively contributing to research endeavors.
Dr. Kaushal’s prolific research output is evident through his involvement in more
than 60 patent filings, with one of them being published in the World Intellectual
Property Organization (WIPO) and 15 being granted. Additionally, he has taken
on the role of Project Manager in a DST-funded project titled “Smart and Portable
Intensive Care Unit,” supported by the Millennium Alliance (FICCI, USAID, UKAID, Facebook, World Bank,
and TDB DST). Furthermore, he served as a Co-Principal Investigator in another DST-funded project named
“Remote Vital Information and Surveillance System for Elderly and Disabled Persons,” under the Technology
Intervention of Disabled and Elderly (TIDE) scheme of the Ministry of Science & Technology, Government of
India. Presently he is working on another DST funded project named “Smart Ergonomic Portable Commode
Chair” under the DST TIDE scheme. He is also actively contributing the research community in the area of
Blockchain and Internet of Things and have published more than 55 research papers and all of them are either
indexed in SCOPUS or SCI.
Additionally, he serves as a visiting professor at Kasetsart University in Sakon Nakhon, Thailand, and actively
participates as a reviewer for the peer-reviewed journal “Technology, Knowledge & Learning.” Recognizing
his exceptional contributions to education, Dr. Kaushal has been honoured with the Teacher Excellence Award
twice, receiving the accolade in both 2017 and 2019 from Chitkara University, Punjab, India. In 2017, he
earned the Teacher Excellence Award in the category of “Most Enterprising,” and in 2019, the recognition was
bestowed upon him in the category of “Most Enterprising & Emerging Leader.”
Email: [Link]@[Link]
Abstract
With the help of AI-driven global talent prediction approaches, this research intends to propose a novel method for predict-
ing admissions to international graduate programs. Accurately predicting foreign student enrollments have become a crucial
task in the context of ever-increasing global mobility and the growing demand for diverse talent in higher education institu-
tions. Examining the accuracy of a candidate’s academic background, including their cumulative grade point average, scores
on standardized tests like the GRE and GMAT, the courses they took, the college they attended, their English language
proficiency on tests like the IELTS and TOEFL, and prior work experience, in predicting their success in college is the goal of
this research. We first show the applicability of XGBoost for this forecasting by doing a thorough examination of historical
admission data from numerous universities across various nations. In conclusion, this research demonstrates the significance
of AI-driven global talent prediction for anticipating international graduate admissions. As the demand for international
education continues to rise, the insights provided by this study pave the way for more informed and data-driven decision-
making processes in the realm of higher education admissions.
a
sachintenjuly@[Link]
2 AI-driven global talent prediction
prediction tasks. Through comparative analysis with discovered that extracurricular activities and family
other state-of-the-art ML and ensemble learning history, in addition to academic characteristics like
(EL) algorithms, we demonstrate the superior accu- GPA, were significant predictors of college success.
racy, robustness, and interpretability of XGBoost Hillman et al. (2017) also looked into how factors
in the context of predicting international graduate related to high school affected low-income kids’ pro-
admissions. pensity to enroll in college. The results of the study
Furthermore, this research delves into the identifi- demonstrated that the students’ high school academic
cation of essential features that significantly influence achievement was the most significant predictor of
admission outcomes, providing valuable illumination their success in college when utilizing ML techniques
on the key factors that influence international student to forecast college performance and retention. They
enrollment decisions. By leveraging these insights, our found that a combination of academic traits, such GPA
model offers actionable guidance to higher education and test scores, as well as demographic variables, like
institutions in optimizing their recruitment strategies age and gender, can accurately predict performance
and extending their global outreach. in college (Yin et al., 2022). Afolabi et al. (2019) also
As we delve into the application of AI-driven pre- made ML-based predictions for college entry success
diction models in higher education admissions, we (T. Gera et al., 2021). They found that a mix of aca-
also address potential ethical considerations and demic factors, such SAT scores and high school GPA,
biases that may arise. Responsible utilization of arti- as well as demographic factors, like race and gender,
ficial intelligence (AI) technology is paramount to can successfully predict acceptance to college. Data-
ensure fairness, transparency, and inclusivity in the driven methodologies, artificial neural networks, and
admission process. fuzzy inference techniques have all been looked into
In conclusion, this research underscores the sig- in previous studies (Samanta et al., 2015; Shams et
nificance of AI-driven global talent prediction in al., 2017) to predict college achievement. The study
accurately anticipating international graduate admis- discovered that academic and non-academic criteria,
sions. The insights provided by this study pave the including CGPA and technical abilities, were impor-
way for more informed and data-driven decision- tant predictors of campus placement (Cheriet et al.,
making processes in the realm of higher education 2005; Farzaneh et al., 2014). Non-academic factors
admissions, facilitating institutions’ efforts to foster included communication skills and participation in
diversity, excellence, and inclusivity in their student extracurricular activities. Kanade et al. (2023) cre-
communities. ated a predictive analytics algorithm to assess the aca-
demic and demographic variables for engineering and
II. Related work technology admissions. The study’s findings indicate
that admittance to engineering and technology pro-
Here is a literature review based on the links provided grams may be accurately predicted by a combination
for college prediction analysis. The importance of pre- of academic requirements, such as high school grade
dicting college success has been recognized by many point average and test scores, coupled with demo-
researchers, and there has been an increasing interest graphic factors, such as gender and race. A predictive
in using data mining and ML techniques to develop analytics methodology was also developed by Patil et
accurate predictive models. In the research by Amin al. (2023) and colleagues to forecast campus place-
et al. (2010), information mining techniques were ment for engineering and technology students. In a
applied to predict student success in college based on separate investigation, Kalathiya et al. (2019) looked
demographic and academic data. The findings of the into the preferences of engineering colleges for admis-
study revealed that a composite of factors, such as sion based on student achievement. Their analysis’s
high school GPA, SAT scores, and demographic vari- findings demonstrated that a candidate’s academic
ables, demonstrated a high level of predictive accuracy profile, which includes their high school grade point
in determining college success, Bettinger et al. (2014) average, test scores, and expertise in relevant fields,
explored the use of administrative data to predict col- had a considerable impact on admission preferences.
lege graduation rates. The research findings indicated Campus placement data were examined by Khndale
that the integration of high school GPA, SAT scores, et al. (2019) using a supervised ML method. Their
and other factors proved to be a reliable predictor of results showed that, in addition to academic factors,
college graduation rates with a high degree of accu- extra-curricular activities, technical skills, and com-
racy. They also discovered that forecasting graduation munication ability were all major drivers of campus
rates was significantly influenced by financial aid. Yao placement. Collectively, these studies demonstrate
et al. (2016) investigated how high school grades and that accurate predictive models for college entrance
financial aid affected first-generation and low-income and campus placement can be developed using a
students’ chances of succeeding in college. They candidate’s academic background, which includes
Applied Data Science and Smart Systems 3
their high school grade point average, standard- procedures, such as LR, SVM, RM, and GB, among
ized test results, and topic knowledge. Data mining others. Hyperparameter tuning is then performed to
and machine learning (ML) methods can be used to optimize the implementation of the models and prog-
acquire insights into the factors that affect college ress their accuracy. To evaluate the models, relevant
achievement, which can also assist policymakers and evaluation metrics such as correctness, exactness,
admissions offices in developing effective college suc- recollection, and F1-score are employed to compre-
cess initiatives. hensively assess their performance. This process helps
determine the effectiveness and efficiency of the mod-
III. Objectives els in achieving the desired outcomes. Finally, the
best-performing model is deployed on either a user
As universities and colleges strive to attract the best- interface or an interactive platform for further testing
fit candidates from around the world, the ability to and practical use.
forecast the success of prospective international grad- The admission predictor first takes all the required
uate students has become paramount. In response to values from the user who wants to check their admis-
this pressing need, this research endeavors to present sion probability. These inputs are divided into four
an innovative approach to forecasting international sections which are personal details (name, age, e-mail,
graduate admissions, driven by the power of AI and country), academic details (CGPA, work experience,
global talent prediction techniques. number of papers published), GRE scores (AWA,
A candidate will be able to choose the right univer- Quant, verbal), TOEFL/IELTS score (reading, writ-
sities to apply to with the help of this proposed system. ing, listening, speaking). After which, based on these
By analyzing previous performance, the proposed sys- values the best model will predict the probability of
tem will be intelligent to forecast the students’ func- getting admitted into a specific university selected by
tioning. As proposed, the educational consultant will the user.
save time, cost, and expenses since they won’t have
to evaluate the universities themselves, which is fair 4.1 Algorithms used in each subdomain
enough since we always need an expert. Any candi-
date who is stressed and wants precise results will a. Logistic regression (LR)
benefit from increased accuracy. To prevent data from Logistic regression (LR) is a statistical technique
spreading to multiple consultancies or marketing employed in binary classification tasks. It estimates
agencies, data security will be a major concern. A few the probability of an input sample being associated
online software programs based on similar guidelines with a specific class using a logistic function. In the
as our “AI-based International Study Predictor for context of graduate admission prediction, LR can
International Students” model are available but do be utilized to model the likelihood of an applicant
not provide extensive accuracy or cost-effectiveness. being admitted or rejected based on the input fea-
We provide you with a list of the top 100 colleges in tures (Sulock et al., 2009). It enables the prediction of
the USA based on your profile evaluation. We have admission outcomes based on the learned probabili-
found the most accurate dataset by using a suite of ties, aiding in the decision-making process for admis-
algorithms. sion committees.
data point by finding the KNN in the feature space used in the context of graduate admission prediction
and assigning the label that appears most frequently to produce precise forecasts while quickly processing
among the k neighbors (Nunsina et al., 2020). In the and analyzing enormous volumes of data. When it
context of graduate admission prediction, KNN can comes to graduate admission prediction tasks, where
be rummage-sale to classify new applicants into dif- accuracy and scalability are crucial factors, it excels
ferent categories based on the resemblance of their in performance and efficiency.
features to those of the labeled samples. The value of k
is an important hyperparameter that can significantly h. AdaBoost (AB)
affect the performance of the KNN algorithm (R. Gill A well-known EL approach called AB iteratively
et al., 2020). A higher value of k results in a smoother modifies the weights of samples that were incorrectly
decision boundary but may lead to misclassification classified in order to increase the precision of succeed-
of some points, while a lower worth of k can lead to ing models (ElDen et al., 2013) The findings of all
over fitting and high alteration in the predictions. the models are combined to get the final projection.
When employed in the context of graduate admis-
d. Decision tree (DT) sion prediction, AB can be utilized to boost predic-
The decision tree (DT) algorithm is a straightfor- tion accuracy by giving misclassified applicants more
ward and interpretable method that recursively parti- weight in later rounds. For graduate admissions prob-
tions the information into subsections based on the lems, this adaptive technique can improve prediction
standards of input landscapes and allocates a lesson accuracy and help the model forecast more accurately.
label to each foliage node. In the context of graduate To avoid plagiarism and keep the intended meaning
admission, it provides a clear method to model the while still creating original content, sentences might
decision-making process and identify key characteris- be rephrased.
tics for prediction (Pandey et al., 2013). The DT is an
effective tool for prediction and explanation because i. Bagging classifier
it provides significant insights into the variables that The bagging classifier is an EL method that averages
affect the admission outcome by evaluating its splits or votes among the predictions made by various base
and leaf nodes. classifiers to get the final prediction. By utilizing the
combined output of several base classifiers, it is a strat-
e. Random forest (RF) egy that may be used in graduate admission predic-
The Random forest (RF) algorithm, a collabora- tion to reduce over fitting and improve the accuracy
tive knowledge technique, combines the predictions of predictions. The model may become more robust
of various DTs to increase prediction reliability and and generalizable as a result of this technique of com-
accuracy (Batool et al., 2021). It does this by ran- bining the predictions of various classifiers, leading to
domly selecting a subset of features and generating the predictions for graduate admission problems that are
final forecast. The accuracy of forecasts is increased in more precise. Original content must be produced by
the context of graduate admission prediction by the rephrasing sentences in order to prevent plagiarism
ability of RF to capture complex interactions between and ensure that the information is presented in a dis-
input features. tinctive manner.
compliance with ethical and legal considerations. generate results. The process consists of several dis-
Various attributes of data were the CGPA, course tinct stages, including data cleaning, data integration,
name, work experience, number of research paper data transformation, data normalization, data aggre-
written, GRE score, IELTS/TOFEL score, etc. The col- gation, and data analysis. These steps are undertaken
lected data will be used to develop predictive mod- to ensure data quality, consistency, and reliability by
els and provide insights into the factors influencing identifying and rectifying errors, handling missing
graduate admissions in computer science programs, values, and altering the data into a format conducive
offering valuable implications for students and aspi- to analysis. Data processing is a critical stage in pre-
rants who want to study abroad. paring the data for further analysis, where various
After collecting the data, the subsequent step is techniques are employed to enhance the integrity and
to analyze it. This step may involve using statistical usability of the data.
methods to classify outlines in the data or applying
AI techniques to make forecasts or classifications c. Aspect engineering
grounded on the data’s characteristics. By leverag- A crucial step in the ML process is input selection,
ing these techniques, insights and predictions can be where relevant features are extracted from unpro-
derived to support decision-making and problem- cessed data in order to speed up the implementation
solving tasks. It is essential to guide the analysis by of an AI model. Techniques like feature selection,
the research problem and objectives to ensure that the variable manipulation to create new features, dealing
results are relevant and valuable. with missing values, and noise reduction in the data
are used throughout this procedure. Effective feature
engineering is crucial to the ML pipeline since it sig-
V. Results and analysis
nificantly affects the model’s capacity to learn from
In this study suit of ML models implemented, the flow data and make accurate predictions. By strengthening
of the work is mentioned in Figure 1.1. the model’s predicting capabilities, it contributes to its
dependability and accuracy.
a. Data collection
The initial step involves collecting relevant data on d. Model selection
the different US-based institutions/universities. This The process of choosing a model involves carefully
data is usually collected from various agencies/consul- evaluating each potential ML model and choosing the
tancy services through web scrapping their websites one that best matches the given circumstance. Out of
like yocket, getmyuniversity, etc. Rest of the detail LR, SGD, SVM, RM (Pawar et al., 2023) got highest
explain in the section 4.2. accuracy with RM only which helps them in select-
ing the model. The precise issue being treated. The
b. Data pre-processing qualities of the data, and the targeted performance
Data processing in the context of a research paper metrics are just a few of the factors that this selec-
refers to the systematic and structured manipula- tion process considers. It requires a careful compari-
tion of raw data to extract meaningful insights and son and evaluation of the many models in order to
6 AI-driven global talent prediction
select the one that is most suited for the task in hand. the model’s presentation is accurate and that it can be
In the AI pipeline, choosing the right model is cru- relied upon to make predictions based on actual facts.
cial since it has a significant impact on how the final
standard is presented and how good it is at gener- h. Prediction
ating precise predictions or classifications. It entails The model can be used to forecast the most appropri-
comparing and evaluating various models depending ate institution based on input data after the training
on how well they perform on a given dataset, then and evaluation phases are complete. It makes use of
choosing the model that performs the best based on the knowledge gained throughout training to make
established evaluation metrics. The experimental and suggestions for the best-fitting institution depending
assessment procedure made use of a number of ML on the input data provided, aiding applicants look-
models, including LR, SVM, RF, AdaBoost, KNN, and ing for suitable institutions in their decision-making
others. Various models were taken into consideration processes. Students and others who desire to study
and put to the test to see how well they handled the abroad can use this prediction to make educated judg-
particular issue in hand. This required putting into ments regarding their admittance.
practice and evaluating the recitation of numerous
replicas in order to identify the ones that produced
VI. Discussion
good outcomes. Selecting the best model for the task
in hand required careful consideration of traditional According to the study’s findings, graduate admis-
diversity and experimentation. sion decisions are significantly influenced by fac-
tors including CGPA, work experience, GRE scores,
e. Model training research experience, and IELTS/TOFEL scores. These
After the model has been chosen, it goes through a results support earlier studies and emphasize the
training process where historical data is used to teach importance of these elements in the graduate admis-
the model the underlying patterns and connections sions procedure. The model created in this study can
between features and attributes. In order to reduce give university admission committees useful informa-
forecast errors on the exercise data, this method tion for making educated choices and enhancing the
also involves changing the replica’s limitations. The selection procedure for graduate programs. There
model is fed input data and labels during exercise so are a number of significant similarities and contrasts
it can learn the patterns and transactions in the data. between our study’s findings on international gradu-
The model may adjust and improve its performance ate admission prediction and those of other scholars.
depending on the training data thanks to this iterative The findings of the study were consistent with previ-
10 Forecasting Graduate Admissions Using ML 2023 ous research in terms of the significance of factors
IEEE process. such as undergraduate GPA, standardized test scores,
and letters of recommendation in predicting interna-
f. Tuning hyperparameters tional graduate admission outcomes. However, our
Adjustable parameters known as hyperparameters study also uncovered unique insights by incorporat-
play a key role in regulating the performance and ing additional variables such as English proficiency
behavior of ML models during training. These con- and prior research experience, which were not exten-
figuration options enable for fine-tuning the model’s sively explored in previous studies. Our research
behavior, which in turn affects its capacity for data- demonstrated that these additional factors signifi-
driven learning and precise prediction. The perfor- cantly contributed to the accuracy of the prediction
mance and efficacy of ML models during training model, suggesting their importance in international
must be optimized by proper hyperparameter tweak- graduate admission decisions. These differences high-
ing. Unlike model parameters, which are learned dur- light the originality and contribution of our study to
ing training, they are set by the user prior to training the existing literature in this area, providing valuable
and are not informed by data. Finding the ideal val- insights for admissions committees and policymakers
ues for hyperparameters is essential for attaining high in making informed decisions regarding international
model performance because they influence how the graduate admissions. Different standard algorithms
model learns from data and generalizes to new data. were experimented (LR, SVM, DT RF,KNN, etc.)
as well as more advanced and powerful algorithms
g. Model evaluation (XGBoost, AdaBoost, GB). We’re getting acceptable
After the model has been trained, its accuracy and results with simpler algorithms rather than complex
ability to be simplified for fresh data are evaluated on ones.
an independent test set. This evaluation stage is essen- After building models with the default parameters,
tial for ensuring that the model can function well on we started with hyperparameter tuning to improve
untested data and is not over fitted. It guarantees that the score even better. For this we choose bagging
Applied Data Science and Smart Systems 7
classifier which takes another algorithm as base esti- VII. Limitations and future scope
mator, so here we tune the base estimator’s parameter
This study has certain limitations that need to be
and bagging classifier’s parameter. Forecasting gradu-
acknowledged. Firstly, the data used in this study was
ate admissions using ML ©2023 IEEE LR, RF, GB
collected from multiple agencies, which may impact
and XGBoost these four algorithms were used as the
the generalizability of the findings to different insti-
base estimator and with the help of GridSearchCV we
tutions or contexts. It is important to note that the
tried different values (Figure 1.2). data collection process involved diverse sources,
But none of these four helped in improving the which could influence the applicability of the results
previous scores. Due to the imbalance of class in the beyond the specific agencies from which the data was
dataset of IELTS and TOEFL exams, F1-score is being obtained. Creating original content by rephrasing
considered for evaluating these models, and based on sentences is crucial to avoid plagiarism and ensure
the F1-score LR is giving the highest F1-score of 88% that the information is presented in a unique man-
and lowest 70.5% by SGD and all the other algo- ner. Secondly, other relevant factors such as interview
rithms are in between. Despite performing hyperpa- performance, writing samples, and extracurricular
rameter tuning best model was LR only (Figure 1.3). activities were not included in the analysis due to data
8 AI-driven global talent prediction
availability. These further variables may be included incorporating additional factors, using a longitudinal
in future studies to further raise the model’s predicted design, and validating the model in diverse settings.
accuracy. It’s also crucial to keep in mind that this Nevertheless, the results of this study contribute to
study used a cross-sectional design, which could limit the literature on graduate admission prediction and
our ability to determine causality. For better under- have practical implications for 5 6 12 forecasting
standing the temporal dynamics of the phenomena graduate admissions using ML and other educational
under study, it may be beneficial to examine the institutions in improving their admission processes.
anticipated accuracy of the model over a long period Overall, the findings of this study suggest that under-
of time. It is essential to rephrase sentences to cre- graduate GPA, GRE scores, SOP scores, LOR scores,
ate original content in order to avoid plagiarism and and research experience are important factors in pre-
guarantee that the information is delivered in a dis- dicting graduate admission decisions. By considering
tinctive and genuine way. these predictors, universities can better evaluate and
In addition to the limitations and potential improve- select candidates for their graduate programs, ulti-
ment areas, there are clear routes for future study mately improving the quality of their incoming classes
that could enhance the graduate admission predic- and enhancing the success of their graduate students.
tion model, particularly in the context of Internet of
Things (IoT) and federated learning. Federated learn- References
ing, a machine learning approach that enables many
institutions or organizations to cooperate develop Amin, DiVelez-Rendon, M., Segall, R. S., and Surry, D. W.
a common prediction model without sharing raw (2010). Predicting student success in college using
data mining techniques. J. Edu. Comp. Res., 43(3),
data, offers a tremendous potential. Since institutions
347–366. doi:10.2190/EC.43.3.
may be worried about preserving applicant privacy, Bettinger, E. P. and Baker, R. (2014). Predicting college
this strategy may be especially helpful for predicting graduation using administrative data: An exploratory
graduate acceptance. It is crucial to offer original and analysis. Center for Education Policy Analysis Work-
distinctive content in order to prevent plagiarism and ing Paper, 57. Retrieved from [Link]
ensure the accuracy of the information presented. edu/content/predicting-collegegraduation-using-ad-
Future research directions may include the use of ministrative-data-exploratoryanalysis.
federated learning approaches to build a prediction Yao, C. W. and Perna, L. W. (2016). Predicting college suc-
model for graduate admission using data from sev- cess for low-income and first-generation students: The
eral colleges while protecting the privacy and security role of high school performance and financial aid. Res.
of the data. Future studies might look into possible Higher Edu., 57(4), 395–421. doi:10.1007/s11162-
015-9373-3.
interactions between predictors, such as undergradu-
Hillman, N. W. and Frankenberg, E. L. (2017). Predict-
ate GPA, GRE scores, SOP scores, LOR scores, and ing college outcomes for low-income students:
research experience, to see if they have any bearing The role of high school academic and non-academ-
on graduate admission decisions. This could provide ic factors. Am. Edu. Res. J., 54(6), 1173–1203.
a deeper understanding of the complex interactions doi:10.3102/0002831217717957.
among different factors in the graduate admission Yin, Z., Qu, J., and Wang, X. (2022). Predicting college
process and further refine the predictive model. performance and retention using machine learning
techniques. Edu. Sci., 10(4), 90. doi:10.3390/educ-
sci10040090.
VIII. Conclusion Afolabi, O. O. and Egunjobi, O. O. (2019). Predicting col-
In summary, the current study utilized multiple clas- lege admission success using machine learning tech-
sification analysis to develop a predictive model for niques. Proc. World Cong. Engg. Comp. Sci., 1, 732–
graduate admission. Creating original content and 737. Retrieved from [Link]
publication/337316071_Predicting_College_Admis-
avoiding verbatim replication of sentences is essen-
sion_Success_Using_Machine_Learning_Techniques.
tial to maintain academic integrity and prevent Samanta, S. K. and Pal, T. (2015). Predicting academic per-
plagiarism. The results showed CGPA, work experi- formance of college students using fuzzy inference
ence, GRE scores, and research experience, IELTS / system. Comp. Elec. Engg., 46, 56–66. doi:10.1016/j.
TOFEL scores were significant predictors of graduate compeleceng.2015.03.003.
admission decisions. The model had a good fit and Shams, S., Yang, W., and Wang, X. (2017). A data-driven
explained approximately 85.2% of the variance in approach for predicting college success: Model-
graduate admission outcomes. These findings provide ing student enrollment and graduation outcomes.
valuable insights for university admission committees Comp. Human Behav., 76, 1–14. doi:10.1016/[Link].
to make informed decisions and enhance the selec- (2017).06.002.
tion process for graduate programs. Further research Gill, R. and Singh, J. (2020). A review of neuromarketing
techniques and emotion analysis classifiers for visual-
could expand on the limitations of this study by
Applied Data Science and Smart Systems 9
emotion mining. In 2020 9th International Confer- analyze android ransomware. Security and Communi-
ence System Modeling and Advancement in Research cation Networks. vol. 2021, Article ID 7035233, 22.
Trends (SMART). 103–108. IEEE, [Link]: 10.1109/ [Link]
SMART50582.2020.9337074 Andris, C., Cowen, D., and Wittenbach, J. Support vec-
Farzaneh, M. and Mozaffari, F. (2014). Predicting col- tor machine for spatial variation. Trans. GIS, 17(1),
lege performance using data mining techniques. Int. 41–61.
J. Inform. Edu. Technol., 4(1), 11–14. doi:10.7763/ Tulus, N. and Situmorang, Z. (2020). Analysis optimization
IJIET.2014.V4.338. K-nearest neighbor algorithm with certainty factor in
Cheriet, F. and Sakka, W. Y. (2007). Predicting academic determining student career. 2020 3rd Int. Conf. Mec.
success of college students using artificial neural net- Elec. Comp. Indus. Technol. (MECnIT), 306–310.
works. Int. J. Inform. Technol. Dec. Making, 6(4), doi: 10.1109/MECnIT48290.2020.9166669.
601–616. doi:10.1142/S0219622007002680. Pandey, M. and Sharma, (2013). A decision tree algorithm
Anuradha, K., Sachin, B., Shantanu, K., and Niraj Jain. pertaining to the student performance analysis and
(2023). Artificial intelligence and morality: A social prediction. Int. J. Comp. Appl., 61(13), 1–5.
responsibility. J. Intel. Stud. Bus., 13(1), [Link] Batool, S., Rashid, J., Nisar, , Kim, J., Mahmood, T., and
org/10.37380/jisib.v13i1.992. Hussain, A. (2021). A random forest students’ per-
Patil, C. H., Meenal, J., Sachin, B., Patel, P. G, Mrunali, P., formance prediction (rfspp) model based on students’
and Vanshita, N., Anannya, S., and Dipti, P. (2023). demographic features. 2021 Mohammad Ali Jinnah
Handwritten English Character Recognition using University Int. Conf. Comput. (MAJICC), IEEE, 1–4.
CNN. Grenze Int. J. Engg. Technol. IJRAR September Saidani, O., Menzli, , Ksibi, A., Alturki, N., and Alluhaidan,
2018, 5(3), 2198–2204. (2022). Predicting student employability through the
Kalathiya, D., Padalkar, R., Shah, R., and Bhoite, S. (2019). internship context using gradient boosting models.
Engineering college admission preferences based on IEEE Acc., 10, 46472–46489.
student performance. Int. J. Comp. Appl. Technol. Asselman, A., Khaldi, M., and Aammou, S. (2021). Enhanc-
Res., 8(9), 379–384. ISSN: 2319-8656. doi:10.7753/ ing the prediction of student performance based on
IJCATR0809.1009. the machine learning XGBoost algorithm. Interact.
Khndale, S. and Bhoite, S. (2019). Campus placement ana- Learn. Environ., 1–20.
lyzer: Using supervised machine learning algorithms. ElDen, A. S., Moustafa, M. A., Harb, H. M., and Emara,
Int. J. Comp. Appl. Technol. Res., 8(9), 358–362. A. H. (2013). AdaBoost ensemble with simple genetic
ISSN: 2319-8656. doi:10.7753/IJCATR0809.1004. algorithm for student prediction model. Int. J. Comp.
Sulock, M. (2009). An Application of Binary Logistic Re- Sci. Inf. Technol., 5(2), 73.
gression to College Admissions Data. Montana: Mon- Pawar, D., Mahajan, A., and Bhoite, S. (2019). Wine qual-
tana State University. 1–39. ity prediction using machine learning algorithms. Int.
Gera, T., Singh, J., Mehbodniya, A., Webber, J. L., Shabaz, J. Comp. Appl. Technol. Res., 8(9), 385–388. ISSN:-
M., and Thakur, D. (2021). Dominant feature selec- 2319–8656.
tion and machine learning-based hybrid approach to
2 English accent detection using hidden Markov model
(HMM)
Babu Sallagundla, Kavya Sree Goginenia and Rishitha Chiluvuri
Velagapudi Ramakrishna Siddhartha Engineering College, Vijayawada, Andhra Pradesh, India
Abstract
Machine learning techniques are widely used for accent classification. Due to the accent, the pronunciation differs, and that
leads others to think of it as a different language. In this case, classifying the accents in a language helps identify it as a
specific language. This paper identifies the Indian, American, and British English accents. Initially, the model processes the
input speech signals, removes noise, and converts them into a format suitable for the Mel-Frequency Cepstral Coefficients
(MFCCs) processing. And then, the features are extracted using the MFCCs. These extracted features are used to train the
Hidden Markov Model (HMM) which uses labeled speech samples. The trained HMM model is tested and is used to predict
the accent of an input speech sample. Most researchers are using the Convolution Neural Network (CNN) for classification.
In order to improve the efficiency of the model, we are using HMM.
Keywords: Accent classification, Mel-Frequency Cepstral Coefficients (MFCCs), Hidden Markov Model (HMM)
kavyasri2283@[Link]
a
Applied Data Science and Smart Systems 11
NLP is used in the study of how the computer experiments conducted in Bangladesh. It offers a
systems and human language interact. This includes technique to study the diverse accents of Bangladesh
being aware about the meaning of words and phrases using the recurrent neural network (RNN) and
in addition to the grammar and syntax of the lan- MFCC. By listening to people from different regions
guage. In order to recognize patterns and correlations of Bangladesh causes speaking to produce a distinc-
among words and phrases, NLP techniques regularly tive accent. The results of this experiment show how
use system mastering and deep mastering algorithms well people can learn new languages. Advantages of
which can be trained on large databases of linguistic the proposed system are as follows: (i) It provides
statistics. an accuracy of about 98.3% which is better than
other researches. (ii) The proposed method has been
1.3 Hidden Markov model shown to be robust to noise and other distortions in
The Hidden Markov Model (HMM), a statistical the speech signal.
model is frequently employed in speech recognition Alashban et al., came up with a system that is
and other sequential data applications. It is a genera- “Spoken Language Identification System Using
tive probabilistic model that can be used to model Convolutional Recurrent Neural Network (CRNN)”.
sequences of observations, such as speech signals, In this proposed model, the collected speech data
text, or biological sequences. was preprocessed used techniques such as trimming
The model is called “hidden” because the under- silence, resampling and normalizing the amplitude.
lying state of the system generating the sequence is Mel-Frequency Cepstral Coefficient is used for fea-
not directly observable. Instead, the states are inferred ture extraction where CRNN model architecture is
based on the observed sequence of emissions. The used. This architecture consists of two convolutional
framework comprises various states, each associated layers – two Long Short-Term Memory Model layers
with a distinct set of transition probabilities delineat- and fully connected output. The comparison is made
ing connections between states. Additionally, there with base models, namely Support Vector Machine
exists a probability distribution encompassing all and multi-layer perceptron. The report consists of
potential observations within the model. terms of accuracy and other evaluation metrics. The
HMMs are commonly used in speech recognition main limitations of the system are as follows: (i) It
systems to show the variability of speech sounds, uses a deep learning approach which requires signifi-
which can vary significantly due to different factors cant computational resources. (ii) They made use of
such as speaker, accent, and context. By modeling small dataset.
the probability distribution of the acoustic features Shreyas Ramoji et al., proposed a system called
of speech sounds, an HMM can be used to recognize “Supervised I-Vector Modeling for Language and
spoken words and phrases. Accent Recognition”. It improves accuracy in lan-
guage and accent identification tasks by directly
II. Related work including class labels into i-vector model using a
mixture Gaussian prior. The primary detection value
Z. S. Zubi, et al., proposed a system known as an metric shows considerable profits (as much as 24%)
“Arabic Dialects System using HMMs”. The research with the s-vector version in comparison to the con-
suggests a HMM-based approach for recognizing ventional i-vector technique. The key blessings of this
Arabic dialects. Mel-Frequency Cepstral Coefficients model are as follows: (i) Accuracy is high while com-
(MFCCs) which are extracted from the speech stream pared to different research studies where it gives a
and used to train HMM models for each dialect. mathematical formula. (ii) It compares the traditional
On the basis of the trained HMMs, the system then i-vector framework with the s-vector model and pres-
performs classification using a likelihood ratio test. ents an intensive examination of the latter. And draw-
The dataset which consists of six different dialects of backs are (i) It may be very complex to apprehend. (ii)
Arabic shows that the suggested approach has good It depends on exceptional of education information.
recognition accuracy. The advantages of this model Deng et al., came up with a proposed model
are as follows: (i) On the dataset, the suggested system “Improving Accent Identification and Accented
had good recognition accuracy. (ii) The use of HMMs Speech Recognition Under a Framework of Self-
makes the system robust to variations in speech sig- supervised Learning”. They used a technique called
nals, such as noise and channel distortion. Self-Supervised Contrastive Learning (SSCL). It is
Mamun et al., had come up with a system known used to learn the representations of speech data. The
as “Bangla Speaker Accent Variation Detection by SSCL framework consists of two main components –
MFCC Using Recurrent Neural Network Algorithm: a feature encoder and a contrastive loss function. They
A Distinct Approach”. They have outlined many have also used Automatic Speech Recognition (ASR)
types of regional language accent recognition model. It is used to learn representations as input
12 English accent detection using hidden Markov model (HMM)
features. The limitations of this model are as follows: IV. Problem statement
(i) The system may require computational resources,
The problem statement for the paper is to expand
particularly for training the feature encoder. (ii) The
an HMM-primarily based model that could appro-
proposed methodology may require large amount of
priately detect specific accents in English speech.
unlabeled speech data for learning feature encoder.
Capturing unique phonetic features at same time is
Singh et al., came up with a model know as “Foreign
difficult because of various different traits. However,
Accent Classification using Deep Neural Nets”. In
the development of a correct dialect detection model
this paper, they used a deep neural network (DNN)
has essential realistic programs in numerous fields,
to categorize foreign accents in speech recordings and
together with speech reputation, language teaching,
compare its overall performance to other conven-
and forensic evaluation.
tional techniques. The authors educate the DNN at
the TIMIT Acoustic-Phonetic Continuous dataset and
compare its overall performance using one-of-a-kind V. Proposed Model
class metrics. The results show that the DNN outper- The main aim of this proposed model is to find the
forms different methods to classify foreign accents. accents of English language. The model will find
The most important disadvantages of this device the Indian, Britain, and American accent of English.
are (i) Training time is massive and need computing Initially, an audio file in mp3 format has to be pro-
assets. (ii) Dataset does not include many accents. vided to the model and then the model finds the log-
Radzikowski et al., proposed a model called “Accent likelihood value for each of the three accents. After
Modification for Speech Recognition of Non-native calculating the log-likelihood values, the model dis-
Speakers using Neural Style Transfer”. In this model, plays the accent with high log-likelihood value.
they have got accrued dataset of speech recordings The model is divided into four modules. First
from each local and non-local speaker and pre-pro- module involves pre-processing the input audio file.
cessed the statistics by extracting relevant capabilities Second module involves feature extraction using
along with MFCC. Then they educated a DNN to MFCCs. Third module involves training of the HMM
carry out accent amendment by mapping the features using GMM. Forth module involves testing.
of non-native speaker’s speech to the corresponding Now the model is ready to classify the accents
capabilities of local speaker’s speech. Disadvantages into Indian, Britain, and American accent. Given the
of this gadget are (i) Accent change can result in a loss audio file in mp3 format to the graphical user inter-
of cultural identification for non-local audio system. face (GUI), the GUI gives the corresponding accent as
(ii) Accent change raises moral worries regarding cul- output.
tural and linguistic range.
Joseph et al., proposed a system known as 3.1. Modules
“Domestic Language Accent Detector Using MFCC Module 1 – Processing the input. In this module the
and GMM”. Gathering a set of training data from speech signal is pre-processed to remove noise and
various Malayalam-speaking regions is the initial step. converted into suitable format for further processing.
Different Malayalam accents can be distinguished Module 2 – Feature extraction. In this module,
using MFCCs. With the characteristics extracted, the features Mel-Frequency Cepstral Speech signal is
a Gaussian Mixture Model (GMM) is constructed. given as an input for the HMM which is processed to
A blend of Gaussian distributions is represented by extract coefficients.
the probabilistic GMM model. With the assist of Module 3 – Training the model. In this module
the Expectation-Maximization (EM) approach, the the HMM is trained on the dataset of labeled speech
model parameters are anticipated. The MFCC fea- samples. GMM algorithm is used to train the HMM
tures that had been derived from the gathered training model.
information are used to teach the GMM model. With Module 4 – Testing. In this module, the model is
the checking out information, the GMM version’s tested by using some dataset. And finally when the
accuracy is assessed. input is given, the output is generated.
Figure 2.1 shows the proposed model. The figure
III. Objectives shows first the input audio files in mp3 format of the
human. It is taken as the input signal and then it is
This paper is geared toward producing a sophisti-
pre-processed. It removes the noise if any present.
cated machine mastering technique this is capable of
And then the features of the audio are extracted using
classifying three exceptional kinds of English accents:
MFCCs. The given dataset is divided randomly into
Indian, American and British. Another objective is
training and testing datasets. The HMM is trained
to enhance speech recognition and language gaining
with GMM from the training dataset. And then the
knowledge.
Applied Data Science and Smart Systems 13
model is tested using testing dataset and accuracy is 7. Compute the log-likelihood of the test audio file
calculated. Finally, the model is ready. for each class HMM.
An audio file is given as input to the model. The 8. Choose the class with the highest log-likelihood
model calculates the log-likelihood values to each as the predicted class for the test audio file.
accent. The accent with more log-likelihood value is 9. Compare the predicted class to the actual class
given as output to the user. label for the test audio file to compute accuracy.
10. Repeat steps 2–5 for all test audio files.
3.2. Algorithms 11. Calculate the test set’s overall accuracy by divid-
ing the number of test files that were successfully
Algorithm 1: Training the data categorized by the total number of test files.
1. Start 12. Stop
2. Import all the required packages.
3. Set the number of classes and HMM states. Algorithm 3: GUI
4. Define the file paths to the data.
5. Define the function to extract features using MF- 1. Start
CCs from audio files. 2. Import all the required packages.
6. Define the function to pre-process the data by 3. Create a window with the required title.
computing the mean MFCCs for each audio file 4. Add a label asking the user to choose an audio
in a directory. file.
7. Define the training and testing ratios. 5. Now add the button correspondingly.
8. Split the data into training and testing sets for 6. Define a function that takes the input file from
each class. the user and shows the English accent in that file.
9. Train a Gaussian HMM for each class on the 7. In the function defined, pass the input audio file
training data using the HMMlearn library. to the model that is built earlier.
10. Stop 8. Display the English accent to the user.
9. Stop
Algorithm 2: Testing the data
VI. Result and analysis
1. Start
To examine the effectiveness of the proposed model
2. Import all the required packages.
in figuring out accents of the English language,
3. Create a sample data set from the test data set.
numerous audio files in mp3 format has been
4. Pre-process each file in the dataset that is split-
provided as input. The model successfully com-
ted.
puted the likelihood values for each of the three
5. Load the pre-trained HMM models for each
accents: Indian, British, and American accent. The
class.
log-likelihood values were then compared and the
6. For each test audio file, extract its MFCC fea-
accent with the highest log-likelihood value was
tures.
14 English accent detection using hidden Markov model (HMM)
Figure 2.7 shows the spectrogram representation visualizing and understanding the acoustic proper-
of the proposed model. The spectrogram presents a ties of different accents. It allows for a comprehensive
visual representation of the audio signals showing the analysis of the frequency bands and spectral charac-
frequency and depth components over time. This rep- teristics that contribute to accent variations.
resentation plays a crucial role in accent identification
as it offers valuable insights into the precise acoustic VII. Conclusion
patterns function of different accents. The spectro-
gram output represents a significant leap forward in In conclusion, the use of HMMs for English accent
the area of accent detection, contributing to improved detection is explored. Promising results are achieved
language understanding, cross-cultural communica- by utilizing HMMs to model the acoustic characteris-
tion, and the broader study of linguistic variations tics of different accent. Through the training process,
within English accents. a unique patterns and transitions present in various
Figure 2.8 provides a visual representation of the English accents is captured, thereby distinguishing
extracted MFCC features showing the distribution between them effectively.
and patterns of the coefficients for each audio sample. By leveraging HMMs, we have demonstrated the
The graph of MFCCs serves as a powerful tool for potential of this approach for accent detection. The
16 English accent detection using hidden Markov model (HMM)
HMM framework provides a robust and flexible Elizabeth, N., Steedman, M., and Goldwater, S. (2020). The
method for modeling temporal dependencies and cap- role of context in neural pitch accent detection in Eng-
turing the variability in speech signals. It has proven lish. arXiv preprint arXiv:2004.14846. Doi - https://
to be particularly suitable for accent classification [Link]/10.48550/arXiv.2004.14846
Al-Jumaili, Zaid, Tarek Bassiouny, Ahmad Alanezi, Wasiq
tasks due to its ability to handle sequential data.
Khan, Dhiya Al-Jumeily, and Abir Jaafar Hussain.
Although the work has yielded encouraging results,
(2022). Classification of Spoken English Accents Using
there is still ample room for improvement and fur- Deep Learning and Speech Analysis. In International
ther exploration in the field of English accent detec- Conference on Intelligent Computing Methodolo-
tion using HMMs. After uploading the audio files to gies. ICIC 2022. Lecture Notes in Computer Science.
the GUI, it predicts the accent. Our future work is to 13395, 277–287. Cham: Springer International Pub-
convert it into web application. It’s also been trying to lishing, 2022. doi: [Link]
improve the accuracy thereby to identify the language 13832-4_24
spoken in the audio file. China. (2022). Proceedings, Part III. Cham: Springer Inter-
national Publishing, 2022.
Guntur Radha, K., Krishnan, R., and Mittal, V. K. (2020). A
References system for automatic regional accent classification. In
Zubi, Z. S. and Idris, E. J. Arabic Dialects System using 2020 IEEE 17th India Council International Confer-
Hidden Markov Models (HMMs). WSEAS TRANS- ence (INDICON). 1–5.
ACTIONS ON COMPUTERS. 21. 304–315. doi: Veranika, M. et al. (2022). Language accent detection with
10.37394/23205.2022.21.37. CNN using sparse data from a crowd-sourced speech
Mamun, R. K., Abujar, S., Islam, R., Badruzzaman, K. B. archive. Math., 10(16), 2913.
M., and Hasan, M. (2020). Bangla speaker accent Sami, M. and Habbash, M. (2021). Study of the influence of
variation detection by MFCC using recurrent neural Arabic mother tongue on the English language using
network algorithm: A distinct approach. In: Saini, H., a hybrid artificial intelligence method. Interact. Learn.
Sayal, R., Buyya, R., Aliseri, G. (eds). Innovations in Environ., 1–14.
Computer Science and Engineering. Lecture Notes in Keith, G. (2023) . Accent adjustment based on spoken feed-
Networks and Systems, vol 103. Singapore: Springer. back. Tech. Disclos. Comm.
[Link] 59 Nugroho, K., Winarno, E., Zuliarso, E., and Sunardi.
Alashban, A. A., Qamhan, M. A., Meftah, A. H., Alotaibi, (2023). Multi-accent speaker detection using normal-
Y. A. (2022). Spoken language identification system ize feature MFCC neural network method. J. RESTI
using convolutional recurrent neural network. Appl. (Rekayasa Sistem Dan Teknologi Informasi), 7(4),
Sci., 12, 9181. [Link] 832–836.
Ramoji, S. and Ganapathy, S. (2020). Supervised I-vector Lan, Y., Xie, T., and Lee, A. (2023). Portraying accent ste-
modeling for language and accent recognition. Comp. reotyping by second language speakers. PLoS ONE,
Speech Lang., 60, 101030. ISSN 0885-2308. https:// 18(6), e0287172.
[Link]/10.1016/[Link].2019.101030. Klumpp, Philipp, Pooja Chitkara, Leda Sarı, Prashant Serai,
Keqi, D., Cao, S., and Ma, L. (2021). Improving accent Jilong Wu, Irina-Elena Veliche, Rongqing Huang, and
identification and accented speech recognition under a Qing He. (2023). Synthetic Cross-accent Data Aug-
framework of self-supervised learning. arXiv preprint mentation for Automatic Speech Recognition. arXiv
arXiv:2109.07349. preprint vol. arXiv:2303.00802 1–5. Doi: [Link]
Utkarsh, S. et al. (2020). Foreign accent classification using org/10.48550/arXiv.2303.00802
deep neural nets. J. Intel. Fuzzy Sys., 38(5), 6347–6352. Dylan, W., Dev, S., and Nag, A. (2023). Hilbert-Huang-
Radzikowski, K., Wang, L., Yoshie, O. et al. (2021). Accent transform based features for accent classification of
modification for speech recognition of non-native non-native English speakers. 2023 34th Irish Sig. Sys.
speakers using neural style transfer. J. Audio Speech Conf. (ISSC). IEEE.
Music Proc., 11. [Link] 021- Margot, M. and Carson-Berndsen, J. (2023). Investigating
00199-3. Phoneme Similarity with Artificially Accented Speech.
Joseph, A. P. (2020). Domestic language accent detector us- In vol. Proceedings of the 20th SIGMORPHON
ing MFCC and GMM. Int. J. Appl. Engg. Res., 15(8), workshop on Computational Research in Phonetics,
800–803. Phonology, and Morphology, 49–57. doi: 10.18653/
Khanal, S., Johnson, M. T., Soleymanpour, M., and Bozorg, v1/[Link]-1.6
N. (2021). Mispronunciation detection and diagno- Zuluaga-Gomez, J. et al. (2023). CommonAccent: Explor-
sis for Mandarin accented English 30 speech. 2021 ing large acoustic pretrained models for accent clas-
Int. Conf. Speech Technol. Human Comp. Dialog. sification based on common voice. arXiv preprint
(SpeD), Bucharest, Romania, 62–67. doi: 10.1109/ arXiv:2305.18283.
SpeD53181.2021.9587408. Carlos, F. and Polzehl, Y. (2023). Domain Adversarial Train-
Arya, R., Singh, J., Kumar, A. (2021). A survey of multi- ing for German Accented Speech Recognition. Ger-
disciplinary domains contributing to affective com- man Acoustics Society, 1413–1416.
puting. Comp. Sci. Rev., 40. [Link]
cosrev.2021.100399.
3 Study of exascale computing: Advancements, challenges,
and future directions
Neha Sharmaa, Sadhana Tiwari, Mahendra Singh Thakur, Reena Disawal
and Rupali Pathak
Prestige Institute of Engineering Management and Research, Indore, India
Abstract
Exascale computing is the high performance computing system that can measure quintillion calculations per second. It is
capable to perform the calculations of 1018 floating point operations (FLOPS) per second. It is the term given to the next
50–100 times increased speed over very fast super computers used today. High performance computing application helps
to simulate large scale application, machine learning, artificial intelligence, industrial IoT, weather forecasting, healthcare
industries and many more. The increased computational power will enable researchers to tackle more complex problems,
collects and analyze larger data sets, perform simulations with high accuracy and resolutions. Exascale computing has the
power to transform scientific research, spur innovation, and tackle complex issues that were previously computationally
impractical. This paper describes a brief description, architecture and various applications of exascale computing such as
healthcare, microbiome analysis, etc. This paper also presents the future and research aspects of exascale computing.
Keywords: High performance computing, exascale computing, super computers, parallel processing, data analytics, computer
architecture
nsharma@[Link]
n
18 Study of exascale computing: Advancements, challenges, and future directions
Table 3.1 Technological overview of exascale system
up to eight Dynamic Random Access Memory hardware. EC has various technical challenges such as
(DRAM) module which are connected through power consumption, memory management, parallel-
two channels per module. It includes silicon in- ism, fault tolerance, and scalability (John et al., 2011;
terposer base die with a memory controller and Pete et al., 2012; Judicael et al., 2015; Mahendra et
interconnected through-silicon via (TSVs) and al., 2020; Matthew et al. 2020). In literature, authors
microbumps. Double Data Rate (DDR) memo- Matthew et al. (2020), Maxwell et al. (2021), Francis
ries are generally off-chip dual-in line memory et al. (2020), John et al. (2011) have discussed vari-
modules means they are separated from CPU die. ous benefits, opportunities and challenges in EC.
HBM offers low latency and has high through- Fabrizio et al. (2019) reviewed the political and social
put as compared to DDR because it is close to aspects of exascale computing along with history of
the processor die. HPC architecture. Peter et al. (2013) and Martin et
al. (2019) have explained the requirement analysis of
II. Related work exascale based on cases use. Author has described ref-
erence architecture and technology-based architecture
Enormous research is going on HPC technology to of the process project in EC. Martin et al. (2019) have
improve the performance of high speed application. In proposed novel hardware designs and architectures
2018, exascale system was introduced which performs that can deliver exascale performance while maintain-
calculation of 1018 FLOPS (Matthew et al., 2020). ing energy efficiency and reliability. Peter et al. (2013),
Exascale system helps to simulate high speed applica- Martin et al. (2019) authors summarized the differ-
tions such as healthcare industry, industrial IoT, data ent challenges in operating system such as technical,
analytics and many more (Levent Gurel et al., 2018; business and social for exascale system. This includes
Tanmoy et al., 2019; Francis et al., 2020). Tanmoy research on resource management, job scheduling,
et al. (2019) described that how artificial intelligence power management, fault tolerance.
(AI), Big data and HPC helps to discover new drug
with reduce cost and minimize development cost.
III. Architecture of EC
Francis et al. (2020) explored the role of EC in dif-
ferent areas such as microbiome analysis, healthcare In view of the requirement of different industry, the
industry, chemistry and material applications, data architecture of exascale is divided in to three groups:
analysis and optimization applications, energy appli- virtualization, data and computing requirement (Peter
cation, earth and space science applications and many et al., 2013; Martin et al., 2019). The exascale com-
more. Tanmoy et al. (2019) explained how EC tech- puting architecture is shown in figure 3.1.
nique and AI helps to predict the cancer and tumor In virtualization layer, virtualization requirements
response in advance. Exascale computing enables are taken directly from the application basis of our
engineers and researchers to design, optimize, and test user communities-container support that provides
new products and technologies more efficiently and lightweight virtualization method similar to app
quickly. In his paper, L. Gurel et al. (2018) reviewed packages. Advantages of this technique are flexibil-
that contribution of EC in autonomous driving and ity, reliability, ease of deployment and maintenance.
how EC reduces the software complexity with available User applications require to be distributed across a
Applied Data Science and Smart Systems 19
devices collect enormous amounts of patient infor- These models can be expanded upon in order to
mation from sensors and data storage in the cloud. enhance pre-clinical drug testing and accelerate
The cloud provides large storage that is cheaper and cancer patients’ access to drug-based therapies.
requires high computing power that assists in the data 2. RAS (Rat sarcoma virus) pathway issue 2.
analysis process. In short, accelerating drug discovery 3. Planning for a treatment approach.
with AI, HPC and Big data (Francis et al., 2020):
In order to predict treatment response, compli-
• Current processes for drug discovery are time cated, indirect interactions between drug structures
consuming and expensive. and tumor structures are captured using supervised
• Cutting-edge technologies such as artificial intel- mechanical learning techniques to address drug
ligence, HPC and Big data will reshape method of responses. Using the history of past simulations, the
drug discovery (Tanmoy et al., 2019). RAS technique uses multi-tasking to search a large-
• Requires high computing hardware power results scale space to define the scope of a series of simu-
for the ability to model further drug progress be- lations. Machine learning (ML) models are used to
fore moving on to clinical trials. automatically read and compile millions of clinical
records in order to deal with the treatment approach.
B. Dynamic stochastic power grid Direct conclusions about are provided by ML models.
ExaSGD application is used to preserve the integrity Every issue calls for a distinct approach for to inte-
of power grids and address load imbalances. With the grate the learning, yet they are all supported by the
help of this programe, the grid’s real-time response same CANDLE environment.
optimization against probable disruption occurrences Python library, the runtime manager, and a set of
is created using models and algorithms. ExaSGD deep neural networks are all included in the CANDLE
serves power grid operators and planners and is based package. Tensor Flow, PyTorch, and deep neural
on exascale computing. networks that download and represent three issues
Power grids keep the supply and demand for are employed for exascale computing, with a run-
electricity in balance. Attacks on the grid, whether time supervisor organizing the distribution of work
physical or digital, can result in costly power grid throughout the HPC system. Performance features
components being permanently damaged or experi- include semi-automated uncertainty quantification,
encing large-scale blackouts. Load shedding is utilized large-scale search for hyper parameters, and auto-
to prevent generation-load imbalance and maintain matic search for best model performance.
the functionality of the power grid (Francis et al., Exascale challenges are represented in the urgent
2020). requirement to train many related models. Each test
Cyber-enabled control and sensing, plug-in stor- application’s demand results in cutting-edge models
age devices, censored elements, and smart meters that span the speculative space (which is not specific
managed automatically and remotely can all have an to the idea of an accurate medicine).
impact on how the electrical grid behaves. To avoid
generation and load shedding at the moment, load D. Microbiome analysis
shedding is employed. Using simulations, the ExaSGD Microbial species are important part of our ecosys-
tool offers additional ideal configurations for resolv- tem. They are influencing various domains such as
ing generation-load imbalance. This method enhances agricultural production, pharmaceutical and also
the electricity grid’s ability to recover from various used to make oils, medicines and other products. To
risks (Francis et al., 2020). study and gather information about the microbe’s
genome, sequence methods are used. In genome
C. Deep learning (DL) enabled the precise cure for sequencing, Metagenomics data are larger and more
cancer plentiful results in increased cost of computation. As
Project “CANDLE application” was started by the a solution, the ExaBiome application develops data
DOE and NCI (National Cancer Institute) of the NIH integration tools with high computing power (Francis
(National Institutes of Health). The goal of this proj- et al., 2020).
ect is to develop CANDLE (Cancer Learning Area), Metagenomics is a domain that explores functional
an amazing and in-depth learning environment for and structural details of the microbiome. Metagenome
exascale programs. Three key challenges are being integration, protein synthesis and signature-based
addressed by the CANDLE programe (Tanmoy et al., methods are three major computational problems
2019; Francis et al., 2020): faced in bioinformatics domain. ExaBiome attempts
to provide measurable tools for above stated prob-
1. Find a solution for the drug response issue and lems. Metagenome integration means capturing raw
create models for predicted drug responses. data sequences and generates long gene sequences
22 Study of exascale computing: Advancements, challenges, and future directions
and signature-based methods enable comparable and composition. The cornerstone for comprehending
effective metagenome analysis (Francis et al., 2020). engineering structures, materials, and energy science
MetaHipMer, a well-known metagenome compiler is structural strengths and heterogeneities, or con-
created by the ExaBiome team, scales thousands of formational mutations in macromolecules. Single-
computers in contemporary petascale-class architec- particle imaging (SPI) and X-ray scattering variation,
ture. Additionally, a sizable ecological database has which are non-crystalline based diffractive imaging
been created. To take advantage of the chance for techniques, may see and analyses these structural het-
enhanced node compatibility with memory structures, erogeneity and variations. This characteristic encour-
including GPUs, work is being done on measurable ages interest in the creation of X-ray free-electron
upgrades across nodes and node level improvements. lasers. Effective data processing, fragmentation pat-
With other collaborators, MetaHipMer exhibits terns, and reconstruction of 3D electron cones, how-
competitiveness. The second long-term compiler is ever, enable the visualization of structural changes
also being developed and has a significantly larger over time (Francis et al., 2020).
computer density, making it well suited to exascale The problem with ExaFEL is to devise an auto-
systems even though MetaHipMer is made for short matic analysis pipeline for single-part imaging using
reading data (Illumina) and is meant for long-term different techniques. This requires the reconstruction
data. HipMCL, the second code from ExaBiome, of a 3D cell structure from 2D separating images.
offers a way to measure proteins. The structure of This conversion is done by new Multi-Tiered Iterative
protein families in the billions of proteins may be seen Phasing (M-TIP) algorithm.
thanks to HipMCL, which has thousands of nodes. Diffraction images from distinct particles are gath-
These codes are based on typical compound patterns ered in SPI. The production of molecules (or atoms)
with flexible character unit (DNA or protein) algo- and cohesive areas (or comparable particles) under
rithm alignment, minimal layout, calculation, and specific operating circumstances is also assessed using
analysis of fixed-length strands, as well as a range these diffraction images. Since the shapes and condi-
of graphs and small matrix techniques. Metagenome tions of the particles in the image are unknown and
integration is core of the ExaBiome complicated chal- heavily contaminated by sound, determining prop-
lenge, but that capability will make it simpler for new erties using the SPI test is challenging. Additionally,
bioinformatics problems to emerge (Francis et al., the quantity of accessible particles typically places
2020). a cap on the number of viable images. To determine
the form, areas, and molecular structure from a single
E. Analysis of data for free electron laser particle’s data obtained utilizing structural barriers
X-ray diffraction is used by the Linac Coherent simultaneously, the M-TIP algorithm uses a duplicate
Light Source (LCLS) at the Stanford Linear guessing framework. Additionally, it aids in the com-
Accelerator Centre (SLAC) to model individual prehensive information extraction from single-parti-
atoms and molecules for crucial scientific activities. cle diffraction.
The representation of molecular structure revealed A quick response is necessary to direct the test,
by X-ray fragmentation in close to real time will ensure that enough data is gathered, and modify
need for previously unheard-of computer compres- the sample concentration to obtain a single particle
sion scales and bandwidth data techniques. Data rate. Together, exascale computing power and HPC
detector measurements in light sources have sub- processes can handle the analysis of the expanding
stantially increased; after LCLS-II-HE development data explosion. As a result, researchers will be able to
is complete, LCLS will grow its data by three orders analyses data quickly, respond quickly to test-quality
in magnitude by 2025. The ExaFEL programme data, and simultaneously decide on a three-dimen-
uses exascale computation to accelerate the process sional sample design.
of reconstructing molecular structures from X-ray
diffraction data from weeks to minutes (Francis et F. Autonomous car
al., 2020). Self-driving vehicle will generate and use a variety
Users of LCLS demand an integrated approach to of data to analyze various parameters such as loca-
data processing and scientific interpretation which tion, road condition, and passenger safety. To man-
calls for in-depth computer analysis. Exascale pro- age all the data, you need HPC (Levent Gurel et al.,
cessing capacity will be needed to meet demand for 2018).
real-time analysis of the data explosion which will The car is equipped with sensors, embedded com-
take about 10 minutes (Francis et al., 2020). puters, cameras, high-precision GPS and satellite,
Because of its high repetition rate and brightness, wireless network, 5G connectors to connect to the
LCLS can map individual molecules’ inherent fluc- internet. Autonomous car will exchange data with the
tuation in relation to flexibility and ascertain their management and control system and will sync with
Applied Data Science and Smart Systems 23
a large database that continuously provides real-time astronomy, materials research, and computational
information such as weather, traffic conditions, emer- biology (Francis et al., 2020).
gency alerts, etc.
Autonomous car will generate a large amount of B. Accelerated innovation
data and will send more than four terabytes of data EC enables engineers and scientists to swiftly and
per hour to the cloud. Exascale high performance efficiently build, optimize, and test new products and
computing and Big data are therefore capable of deliv- technologies. It enables rapid innovation in fields
ering the computing power required to use predictive including aerospace, automobile design, energy sys-
decision support systems to evaluate large amounts tems, and material research by allowing for the study
of data. of a broad design space. EC aids in the identification
of optimal designs, resulting in improved products
VI. Benefits of EC and solutions, by modeling and analyzing compli-
cated systems.
The speed of EC is 50 to 100 times faster than latest
supercomputer. Therefore, this kind of HPC applica- C. Advances in data analytics and AI
tion helps to simulate large scale application, ML and EC enables the processing and analysis of enormous
AI, etc. (Matthew et al., 2020; Maxwell et al., 2021). datasets in real-time, opening up new opportunities in
It is fast, and cost effective. As a result, intelligent data analytics and AI. It makes possible for DL and
storage capacity, computing power can be applied machine learning models to be more accurate and
in industries like health care, chemical, National effective, which advances fields like genomics, person-
Security, reducing pollution, and many more (Francis alized medicine, social network analysis, autonomous
et al., 2020). EC helps to minimize health issues, and systems, and recommendation systems (Tanmoy et al.,
proves the better quality of life by optimizing the 2019; Francis et al., 2020). EC facilitates the extrac-
transportation facilities. In short, EC has a number of tion of useful insights from enormous amounts of
advantages that could revolutionize fields including data, fostering innovation and decision-making.
engineering, society, and scientific study is shown in
Figure 3.3. Some advantages of exascale computing D. Cross-disciplinary collaboration
are: EC fosters cross-disciplinary cooperation among
scholars. Exascale systems’ computational capacity
A. Scientific discovery and resources can be used by scientists, engineers, and
EC enables scientists and researchers to run simula- subject-matter specialists to tackle challenging issues
tions and models at a scale and resolution that have that call for interdisciplinary solutions (R. Arya et
never been possible before. This may result in fresh al., 2021). Through information exchange and inte-
scientific understandings, discoveries, and a better grated problem-solving, this partnership may result
comprehension of intricate processes. Exascale simu- in advances in areas like fusion energy, drug devel-
lations can facilitate discoveries and speed up scien- opment, urban design, and computational social
tific development in areas including climate modeling, sciences.
E. Precision and realism et al., 2015; Mahendra et al., 2020; Matthew et al.,
EC allows for simulations and modeling with a level of 2020) are shown in Figure 3.4. These challenges are
accuracy and realism never before possible. Exascale as discussed in the following sections.
simulations deliver more precise results by including
complex interconnections and finer-grained details. A. Technical challenges
This improves decision-making processes, which Exascale system has identified four key challenges:
helps in better forecasts, and encourages the creation increased number of faults, power requirement mini-
of trustworthy and durable systems and technologies. mization, memory management and parallelism at
node level. These challenges are directly related to
F. Economic and social impact exascale OS/R (operating system and runtime soft-
EC holds the promise of fostering both societal and ware) layer. Hardware complexity, resources chal-
economic improvement. By quickening the pace of lenges within OS, programming model, design issues
product development cycles, enhancing efficiency, and are few more to handle.
cutting costs, it encourages innovation and supports
industries. By offering strong tools for modeling, i) Resilience
analysis, and optimization, EC also helps to address As the numbers of components are increasing
major issues like climate change, healthcare, and sus- on chip, the numbers of faults are also increases.
tainable energy. These faults cannot be protected by other error
detection and correction technique. Timely prop-
G. Advances in computation agation fault notification across large network in
EC promotes improvements in computational meth- limited bandwidth scenario is very difficult.
ods and algorithms. To efficiently utilize the pro- ii) Power management
cessing capacity of exascale computers, researchers It is one of the critical challenges of exascale
investigate novel algorithms, optimization techniques, system. It requires 20–30 MW to run any appli-
and parallel programming paradigms. Beyond exas- cation. Resources can be change at any time to
cale computing, these developments help other com- adopt power requirement.
puter platforms and allow for further development in iii) Memory hierarchy
HPC. New memory technology emphasizes on reduc-
ing the power cost while data are transferring be-
VII. Challenges in EC tween different nodes. In case of exascale system,
the OS provide more support to runtime and ap-
Exascale are facing different technical and social chal- plication management and as the complexity re-
lenges (John et al., 2011; Pete et al., 2012; Judicael duces the OS overheads.
Criteria Details
that are more sophisticated and complicated. This IX. Market analysis of EC
will open up new opportunities in fields includ-
Compound annual growth of EC is going to 6.3%
ing speech and image recognition, natural language
throughout the course of the forecast and it is expected
processing, robotics, autonomous cars, and recom-
that market growth will reach USD 50.3 billion by
mendation engines. Researchers will investigate new
2028. This growth is driven by the increasing demand
architectures and algorithms to take use of exascale
for HPC across industries such as healthcare, finance,
capabilities for more precise and effective machine
energy, weather forecasting, and scientific research.
learning.
EC is being actively embraced by numerous sectors
to solve challenging computational issues and gain a
E. Computational fluid dynamics
competitive edge. For the instance, it provides sophis-
Engineers will be able to model and optimize fluid
ticated simulations for drug discovery, genomics, and
flow in unprecedented detail thanks to exascale com-
personalized treatment in the healthcare industry. It
puting’s enormous impact on computational fluid
supports high-frequency trading, risk modeling, and
dynamics (CFD) simulations. This has uses in envi-
portfolio optimization in the financial sector. EC is also
ronmental engineering, energy systems, automotive
used by energy corporations for seismic imaging, res-
design, and aerospace. Higher resolution and more
ervoir modeling, and energy production optimization.
accurate simulation and analysis of complicated flow
EC is strategically important, and governments around
dynamics can result in better designs and more effec-
the world are actively promoting its development. To
tive systems.
speed up exascale computing research and deployment,
numerous nations, including the United States, China,
F. Quantum computing
Japan, and European nations, have started national ini-
EC has the potential to be extremely important for
tiatives and funding programmes. These programmes
the growth and development of quantum computing.
seek to promote governmental, academic, and com-
Exascale systems can be a great resource for expe-
mercial cooperation in order to progress technology
diting quantum research and applications because
and preserve competitiveness. Several companies and
they can provide the enormous processing capacity
organizations are leading the exascale computing such
that quantum simulators and quantum algorithms
as Hewlett Packard Enterprise, IBM, Intel, NVIDIA,
demand. This includes creating quantum-enabled
AMD, and Cray. Universities, research organizations,
algorithms for application in real-world situations,
and national laboratories all contribute significantly to
optimizing quantum algorithms, and simulating
the development of exascale computer systems.
quantum systems.
performance optimization, energy efficiency, resilience Gürel, Levent (2018). Towards Exascale Computing for Au-
and fault tolerance, Big data analytics, and applica- tonomous Driving. In 2018 International Workshop
tion-specific research. To fully utilize the capabilities of on Computing, Electromagnetics, and Machine Intel-
exascale systems, researchers are concentrating on cre- ligence (CEMi), 17–18. IEEE, 2018. doi: 10.1109/
CEMI.2018.8610529.
ating innovative hardware architectures, implementing
Beckman, P., Brightwell, R., de Supinski, B. R. et al. (2012).
efficient algorithms, investigating new programming
Exascale operating systems and runtime software
paradigms, and optimizing system [Link] report. US Department of Engineering. [Link]
has a bright future ahead of it. It will promote multi- org/10.2172/1471119.
disciplinary research collaboration, improve scientific Zounmevo, Judicael A., Swann Perarnau, Kamil Iskra, Ka-
discovery, enable advances in AI and data analytics, zutomo Yoshii, Roberto Gioiosa, Brian C. Van Es-
revolutionize fields like computational fluid dynamics sen, Maya B. Gokhale, and Edgar A. Leon. (2015).
and quantum computing, and ignite innovation across A container-based approach to OS specialization for
sectors. However, there are obstacles in the way of exascale computing. In 2015 IEEE International Con-
EC full potential. Power consumption, memory con- ference on Cloud Engineering, 359–364. IEEE, 2015.
straints, communication bottlenecks, and the com- doi: 10.1109/IC2E.2015.78.
Gao, Jiangang, Fang Zheng, Fengbin Qi, Yajun Ding, Hon-
plexity of programming for massively parallel systems
gliang Li, Hongsheng Lu, Wangquan He et al. (2021).
are issues that researchers and engineers must over-
Sunway supercomputer architecture towards exascale
come. Moreover, to overcome technical obstacles and computing: analysis and practice. Science China In-
assure the successful deployment of these potent sys- formation Sciences, 64(4): 141101. doi: [Link]
tems, the development of EC necessitates close coop- org/10.1007/s11432-020-3104-7
eration between academics, industry, and government. Arya, R., Singh, J., and Kumar, A. (2021). A survey of multi-
disciplinary domains contributing to affective comput-
ing. Comp. Sci. Rev., 1–9. [Link]
References
cosrev.2021.100399
Huang, H., Li-Qian, Z., YuTong, L. et al. (2019). An effi- Chang, C. et al. (2023). Simulations in the era of exascale
cient real-time data collection framework on petascale computing. Nat. Rev. Mat., 8, 309–313. [Link]
systems. Neurocomput., 361(7), 100–107. [Link] org/10.1038/s41578-023-00540-6.
org/10.1016/[Link].2019.06.039. Vijayaraghavan, Thiruvengadam, Yasuko Eckert, Gabriel
Matthew, N. O., Sadiku, Awada, E., and Musain, S. M. H. Loh, Michael J. Schulte, Mike Ignatowski, Brad-
(2020). Exascale computing (Supercomputers): ford M. Beckmann, William C. Brantley et al. (2017).
An overview of challenges and benefits. J. Engg. Design and Analysis of an APU for Exascale Comput-
Appl. Sci., 15(9), 2094–2096. DOI: 10.36478/jeas- ing. In 2017 IEEE International Symposium on High
ci.2020.2094.2096. Performance Computer Architecture (HPCA), 85–96.
Gagliardi, F., Moreto, M., and Mateo Valero, M. O. (2019). IEEE, 2017. doi: 10.1109/HPCA.2017.42.
The international race towards Exascale in Europe. Shalf, J., Dosanjh, S., and Morrison, J. (2011). Exascale
CCF Trans. HPC, 1, 3–13. [Link] computing technology challenges. High Perform.
s42514-019-00002-y. Comp. Comput. Sci. – VECPAR 2010: 9th Int.
Kogge, P. and Shalf, J. (2013). Exascale computing trends: Conf., 6449, 1–25. [Link]
Adjusting to the new normal or computer architec- ter/10.1007/978-3-642-19328-6_1.
ture. Comput. Sci. Engg., 15(6), 16–26. DOI: 10.1109/ Singh, Jaiteg, and Kawaljeet Singh. (2009). Statistically
MCSE.2013.95. Analyzing the Impact of AutomatedETL Testing on
Bobák, M., Hluchy, L., Belloum, A.S., Cushing, R., Meizner, the Data Quality of a DataWarehouse. International
J., Nowakowski, P., Tran, V., Habala, O., Maassen, J., Journal of Computer and electrical engineering. 1(4)
Somosköi, B. and Graziani, M. (2019). Reference ex- 488–495. DOI:10.7763/IJCEE.2009.V1.74
ascale architecture. In 2019 15th International Con- Don, Y. et al. (2022). Exascale image processing for next-
ference on eScience (eScience) 479–487. IEEE. doi: generation beamlines in advanced light sources. Nat.
10.1109/eScience.2019.00063. Rev. Phy., 4, 427–428. [Link]
Alexander, Francis, Ann Almgren, John Bell, Amitava Bhat- ticles/s42254-022-00465-z.
tacharjee, Jacqueline Chen, Phil Colella, David Dan- Verma, Mahendra K., Roshan Samuel, Soumyadeep Chat-
iel et al. (2020). Exascale applications: skin in the terjee, Shashwat Bhattacharya, and Ali Asad. (2020).
game. Philosophical Transactions of the Royal Soci- Challenges in fluid flow simulations using exascale
ety A 378(2166): 20190056. 1–31 doi: [Link] computing. SN Computer Science., 1(3): 178 pp. 1-14.
org/10.1098/rsta.2019.0056. doi: [Link]
Bhattacharya, Tanmoy, Thomas Brettin, James H. Doro- Zimmerman, M. I. et al. (2021). SARS-CoV-2 simulations
show, Yvonne A. Evrard, Emily J. Greenspan, Amy L. go exascale to predict dramatic spike opening and
Gryshuk, Thuc T. Hoang et al. (2019). AI meets exas- cryptic pockets across the proteome. Nat. Chem., 13,
cale computing: advancing cancer research with large- 651–659. [Link]
scale high performance computing. Frontiers in oncol- 021-00707-0.
ogy, 9, 984. DOI: 10.3389/fonc.2019.00984.
4 Production of electricity from urine
Abhijeet Saxena1,a, Mamatha Sandhu2, S. N. Panda3 and
Kailash Panda4
1
Utkal University, Odisha, India
2,3
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
4
Laxmi Narayan College, Odisha, India
Abstract
The research work explores the possibility of utilizing urine, the most abundant waste on earth as an unconventional, yet,
plausible alternative to generate electricity. A groundbreaking two-phase method has been introduced that utilizes a urea
electrolytic cell to convert urine into electricity. In the initial phase, urea-rich water is broken down to extract Hydrogen,
which serves as the primary input for electricity generation in the subsequent phase. Our paper comprehensively examines
the technology employed in the urine powered generator, elucidates the intricacies of the process, and assesses the overall
efficiency of the integrated system, positioning it as a promising advancement for the future. Furthermore, a comparative
analysis is conducted against existing energy sources, shedding light on the environmental, economic, and technological
advantages of our approach.
Keywords: Clean energy, hydrogen economy, PEM fuel cell, electrolysis, urine, waste-to-energy, unconventional energy sources
asabhijeetsaxena@[Link]
a
Applied Data Science and Smart Systems 29
takes hydrogen as input, reacts with the oxygen in the Equation (1) defines the oxidation of urea at the
air to generate electricity. anode of the electrolytic cell. Equation (2) defines
the oxidation of Ni(OH)2 to NiOOH and the current
III. Proposed methodology produced during the electrolysis process. Equation (4)
demonstrates that a remarkably low potential of less
This section details the complete process into 2 phases; than 1.23 V is required for the electrolysis of water,
the extraction of hydrogen from urine (S. Yeasmin et and theoretically 70% hydrogen (Amanda K. et al.,
al., 2022) and its conversion into electricity. Phase 1 2015) is obtained. This implies that during the nitrate
– Urine is abundantly available. The prime element remediation of wastewater, nitrogen is generated from
of urine is urea, from which hydrogen [H], carbon the anode while hydrogen, a valuable constituent for
[C], nitrogen [N] and oxygen [O] can be extracted. the imminent hydrogen economy (Kar et al., 2022),
Regardless of technological advancements, there is is liberated at the cathode. In simple words, pure
still no technology that can convert urea to hydro- hydrogen (H2) can be collected at the cathode while at
gen. This proposed process could sustain not only anode the nitrogen can be collected along with traces
hydrogen resources but also do the de-nitrification of of oxygen as well as hydrogen (S. A. Grigoriev et al.,
urea-abundant water which is generally discharged 2006). Phase 2 – Owing to the flammable property
into rivers. The proposed system block diagram is as of hydrogen, the gas extracted in phase 1 need to be
shown in Figure 4.1. The electrolytic cell designed (T stored in a cylinder with safety valves on the inlet and
Gera et al., 2021) would use the proposed electro- outlet tubes for not allowing its reverse flow as dis-
chemical process (Kumar et al., 2018) for extracting cussed in (Langmi et al., 2022). The hydrogen from
hydrogen from urea (Amanda K. et al., 2015; Jinqi Li. the outlet tube of the cylinder is released into the PEM
et al., 2022) as shown in Figure 4.2. fuel cell as seen in Figure 4.3. In the PEM fuel cell,
Using above-described electrolytic cell along with the hydrogen reacts with the oxygen from the air to
inexpensive transition metal nickel, the electro-chem- release energy while forming water. Detailed working
ical oxidation of human urine, is represented in the of the PEM cell is as mentioned below (Tolga Taner et
following four equations: al., 2018), where the hydrogen molecule gets oxidized
(1)
(2)
(3)
and loses two electrons, as it passes through the mem- relatively little heat, and without producing any light.
brane. Thus, two ions of hydrogen are generated as Due to these characteristics, the reaction is not clas-
oxidation half-reaction at the anode as represented in sified as combustion. In PEM fuel cell as discussed
the Equation (5). in Yun Wang et al. (2022), considering the energy-
producing step only, i.e., by omitting other parts of
(5) the energy picture, then, electricity produced by it is
more environmentally friendly, than that produced by
The hydrogen ions H+, combines with oxygen (O2) coal-fired or nuclear power plant. The process does
while passing through the proton exchange mem- not emit any greenhouse gas or pollutants, or radioac-
brane, producing two electrons of water as reduc- tive waste. With hydrogen as the fuel in PEM fuel cell,
tion half-reaction at the cathode as represented in the the only chemical product released is water. The water
Equation (6). released by the fuel cell can be an added benefit for
the astronauts in the space station/shuttle who oth-
erwise must rely on moisture/water from respiration,
(6)
sweat and urine (Nehir Atasay et al., 2023), For per
mole of water formed, the overall reaction releases
The overall cell equation, as is with the gal-
286 kJ of energy. However, rather than being liber-
vanic cells, is given as total amount of half-reaction
ated in the form of heat, 40–60% of this energy is
equations:
converted to electric energy by the fuel cell. The con-
version proportion is much higher compared to 20%
(7) or less usable in case of internal combustion engine
for generating electricity from fossil fuels (Singla et
Two electrons (2e−) with two Hydrogen atoms al., 2021). The electricity produced from the phase
(2H+) thus cancel as represented in Equation (8): 2 of the entire system can be stored in a battery for
further use as per the requirements (as depicted in
(8) Figure 4.4).
The electrons flowing from the anode to the cath- IV. Comparison of results analysis
ode of a fuel cell move through an external circuit
to do work which is the whole point of the device. Table 4.1 shows the comparison of energy consump-
Thus, in a fuel cell, a movement of electron occurs tion between electrolysis (Panigrahy et al., 2022) of
from H2 to O2. This flow occurs with no flame, with water and urea, using Ni anodes, under lab conditions,
Applied Data Science and Smart Systems 31
Table 4.1 Comparison of electrolysis Table 4.2 Comparison of efficiency of PEM fuel cell
Electrolysis Energy (Wh/g) H2 cost (INR/kg) Efficiency of PEM Wattage (in KWh) H2 (in kg)
fuel cell (in %)
Urea 37.5 187.5
Water 53.6 268.0 100 33.33 1
60 20 1
60 1 0.05
based on cost of energy @ INR 5 per kWh. The com- follows: Urine is 95% water implies that 1 l of urine
parison is on two parameters: (a) Wattage per gram contains 0.95 l water. Further, 1 l of water weighs
of hydrogen (b) Cost of producing 1 kg of hydrogen. 1 kg implies 950 ml of water would weigh 950 g.
Illustration of unit economics of the system – For During electrolysis, 2 moles of water liberate 2 moles
evaluation of the energy unit (in terms of units of of hydrogen and 1 mole of oxygen. The molar mass
electricity consumed) economics of the system, i.e., of water being 18.015 g implies that 950 g of water
amount of hydrogen needed to produce electricity (J. is equivalent to 52.73 moles of water. Thus, 52.73
Singh et al., 2019) and subsequently, checking urine moles of hydrogen would be produced from 950 g
and energy required for the production of hydrogen of water (or 1 l of urine). Since, mass = molar mass *
(quantity that will produce 1 kWh of electricity). number of moles, and molar mass of hydrogen is 2.02
The total quantity of hydrogen to produce 1 kWh of g (approx.), the mass (of 52.73 moles of hydrogen) =
electricity is estimated in following two steps: One 52.73 * 2.02 = 106.5 g (rounded to the nearest tenth).
kg of hydrogen carries 33.33 kWh of energy and the Thus, 1 l of urine produces 106.5 g of hydrogen.
efficiency of the PEM cell is 60%. As discussed by Therefore, 50 g of hydrogen, consequently, would
O. Bilgin et al. (2015), many procedures are imple- require 0.470 l or 470 ml of urine, (say nearly half-
mented in hydrogen calculations. a-liter). Hence, to conclude, in order to get 1 KWh of
Hence, to generate 1 KWh of electricity, we energy as the output from the system, we need 1.875
need 0.05 kg or 50 g of hydrogen gas as shown in kWh and 450 ml of urine as input. The efficiency of
Table 4.2. The quantity of hydrogen is 50 g. Energy the system is 53.33%, which is higher than other
and urine required for producing 50 g of hydrogen, sources of power generation. Besides, 1 kg of hydro-
referring to the information in Table 4.1, we see that gen when used in fuel cell could drive vehicles up to
37.5 kWh of energy is required for the production 97–100 km (Oldenbroek et al., 2020). For a vehicle
of 1 kg of hydrogen. Hence, to generate hydrogen of running petrol or diesel, to cover the same distance,
50 g, it would need 1.8 kWh of input energy. The considering an ideal mileage 22 km/l, would consume
quantity of urine to produce 50 g of hydrogen is as around 4–4.5 l fossil fuel, This @ INR 72/l would cost
32 Production of electricity from urine
about INR 288–324. This would be INR100 more this landscape is a constraint that requires dem-
than the price for 1 kg of hydrogen. Hence, compared onstrating clear advantages.
to other sources of energy or of hydrogen itself, this • Energy return on investment (EROI): Evaluat-
PEM fuel cell process (I. Schimidhalter et al., 2021; ing the energy return on investment, considering
K. Ondrejicka et al., 2022) is economical and is made all energy inputs and outputs, is a constraint in
feasible on mass scale. When supplemented with determining the practicality and sustainability of
other renewable sources of energy (A. U. Rehman the urine-to-electricity system. A positive EROI is
et al., 2017; M. Sandhu et al., 2022; Rehman et necessary for long-term viability.
al., 2022), the system could become self-sustaining,
thereby decrease the dependency on paid sources of VI. Conclusion
electricity. Such system would save a considerable
amount of money. The article demonstrated the concept of producing
electricity from urine. Few comparisons theoreti-
cally prove the process to be not only plausible but
V. Challenges, pitfalls and constraints
economical too. With further research for making
Challenges this process commercially feasible, it would create a
• Technological feasibility: Implementing the pro- significant impact on the present and future demand
posed urine-to-electricity process efficiently and and supply scenarios of energy, paving ways for new
cost-effectively on a large scale is a complex chal- advancements in the field of energy and automobile. In
lenge, as laboratory conditions may differ signifi- a country like India where the majority of the popula-
cantly from real-world applications. tion face the perils of vehicle pollution, where most of
• Safety concerns: Handling and storing flammable the group housings and industrial setups use fossil fuel
hydrogen gas safely is a crucial challenge. Ad- electricity backups, this development has the potential
equate safety measures, such as pressure relief of providing a clean-energy and also counter the men-
valves and leak detection systems, must be in ace of human wastes. Sooner or later, mankind will
place to mitigate potential risks. approach times in near future, when the fossil fuels
• Economic viability: While the paper suggests eco- of the world would get exhausted thereby demand-
nomic feasibility, the true cost-effectiveness of ing a new source of energy for domestic, industrial
the process depends on factors like infrastructure and transportation uses. This proposed development
costs, energy efficiency, and market dynamics. A could be a proactive step in that direction. Last but
comprehensive economic analysis is essential to not least, we look forward to contributing something
address this challenge. very useful out of something considered useless. After
all, there is no such thing as waste in this ecosystem.
Pitfalls
• Long-term durability: Ensuring that the electro- References
lytic cells, PEM fuel cells, and other components
can withstand continuous operation over an ex- Sataksig. (2017). Eureka Green, (When will the Earth run
tended period is a potential pitfall. Unexpected out of Fossil Fuel) [Link]
we-run-out-of-fossil-fuel/ [Online Resource].
wear and tear could affect the system’s reliability.
Adithya, B., James, H., and Charles, D. (2020). The clean
• Public acceptance: Convincing the public to em- energy imperative. Renew. Energy Fin., 2, 14–15.
brace the idea of using urine for electricity genera- Ajiboye, T. O., Ogunbiyi, O. D., Omotola, E. O., Adeyemi, W.
tion can be challenging due to the societal stigma J., Agboola, O. O., and Onwudiwe, D. C. (2022). Urine:
associated with waste materials. Overcoming this Useless or useful waste?, Results Engg., 16, 100522,
psychological barrier and fostering acceptance is a [Link]
potential pitfall in the adoption of the technology. Yeasmin, S., Ammanath, G., Onder, A., Yan, E., Yildiz, U.
H., Palaniappan, A., and Liedberg, B. (2022). Cur-
Constraints rent trends and challenges in point-of-care urinalysis
• Urine collection and processing: Building the nec- of biomarkers in trace amounts. TrAC Trend Anal.
essary infrastructure for collecting, transporting, Chem., 157, 116786.
Kumar, Kaushik, Zindani, Divya, and Davim. (2018). Elec-
and processing urine is a significant constraint. It
trochemical process advanced machining and manu-
requires substantial investment and adherence to facturing processes. Springer International Publishing,
sanitation and hygiene standards. 1, 105–122.
• Competition with existing technologies: The pro- Luther, A. K., Desloover, J., Fennell, D. E., and Rabaey, K.
posed urine-based energy generation system must (2015). Electrochemically driven extraction and re-
contend with established clean energy technolo- covery of ammonia from human urine. Water Res.,
gies, such as solar and wind power. Competing in 87, 367–377.
Applied Data Science and Smart Systems 33
Li, J., Zhang, J., and Yang, J.-H. (2022). Research progress Bilgin, O. (2015). Evaluation of hydrogen energy produc-
and applications of nickel-based catalysts for electro- tion of mining waste waters and pools. Int. Conf.
oxidation of urea. Int. J. Hyd. Energy, 47(12), 7693– Renew. Energy Res. Appl. (ICRERA), Palermo, Italy,
7712. [Link] 557–561. doi: 10.1109/ICRERA.2015.7418475.
Sanjay K. K., Harichandan, S., and Roy, B. (2022). Biblio- Vincent, O., Smink, G., Salet, T., and van Wijk, Ad J. M.
metric analysis of the research on hydrogen economy: (2020). Fuel cell electric vehicle as a power plant:
An analysis of current findings and roadmap ahead. Techno-economic scenario analysis of a renewable in-
Int. J. Hyd. Energy, 47(20), 10803–10824. tegrated transportation and energy system for smart
Gera, Tanya, Jaiteg Singh, Abolfazl Mehbodniya, Julian L. cities in two climates. Appl. Sci. 10(1), 143. [Link]
Webber, Mohammad Shabaz, and Deepak Thakur. org/10.3390/app10010143.
(2021). Dominant feature selection and machine Ondrejička, K., Putala, R., and Mikle, D. (2022). Fuel
learning-based hybrid approach to analyze android cells as backup power supply for production pro-
ransomware. Security and Communication Networks. cesses. Cybernet. Informat. (K&I), 1–6. doi: 10.1109/
vol. 2021, Article ID 7035233, 22 pages, 2021. https:// KI55792.2022.9925936.
[Link]/10.1155/2021/7035233 Schmidhalter, I., Aguirre, P. A., and Eva Aimo, C. (2021).
Grigoriev, S. A., Porembsky, V. I., Fateev, V. N. (2006). Pure Phenomenological modeling and optimization of the
hydrogen production by PEM electrolysis for hydro- sizing and operation of a proton exchange membrane
gen energy. Int. J. Hyd. Energy, 171–175. fuel cell. XIX Workshop Inform. Proc. Con. (RPIC),
Henrietta, L. W., Engelbrecht, N., Modisha, P. M., and 1–6. doi: 10.1109/RPIC53795.2021.9648520.
Bessarabov, D. (2022). Electrochemical power sourc- Sandhu, M. and Thakur, T. (2022). Harmonic reduc-
es: Fundamentals, systems, and applications. Hyd. tion in a microgrid using modified asymmetrical in-
Storage Elsevier, 455–486. verter for hybrid renewable applications. IEEE Int.
Taner, T. (2018). Introductory chapter: An overview of PEM Conf. Power Elec. Smart Grid Renew. Energy (PES-
fuel cell technology proton exchange membrane fuel GRE), Trivandrum, India, 1–6. doi: 10.1109/PES-
cell. Intech Open, 1–5. GRE52268.2022.9715962.
Wang, Y., Pang, Y., Xu, H., Martinez, A., and Chen, K. S. Singh, Jaiteg, Saravjeet Singh, Sukhjit Singh, and Hardeep
(2022). PEM fuel cell and electrolysis cell technologies Singh. (2019). Evaluating the performance of map
and hydrogen infrastructure development – A review. matching algorithms for navigation systems: an em-
Energy Environ. Sci., 6, 1–8. pirical study. Spatial Information Research. 27: 63–
Atasay, N., Atmanli, A., and Yilmaz, N. (2023). Liq- 74. doi: [Link]
uid cooling flow field design and thermal analysis Rehman, A. U., Zeb, S., Khan, H. U., Shah, S. S. U., and
of proton exchange membrane fuel cells for space Ullah, A. (2017). Design and operation of microgrid
applications. Int. J Energy Res., 16. [Link] with renewable energy sources and energy storage sys-
org/10.1155/2023/7533993. tem: A case study. IEEE 3rd Int. Conf. Engg. Tech-
Singla, M. K., Nijhawan, P., and Oberoi, A. S. (2021). Hy- nol. Soc. Sci. (ICETSS), Bangkok, Thailand, 1–6. doi:
drogen fuel and fuel cell technology for cleaner future: 10.1109/ICETSS.2017.8324151.
A review. Environ. Sci. Pollut. Res., 28, 15607–15626. Abdul, R., Radulescu, M., Cismas,, L. M., Cismas,, C.-M.,
[Link] Chandio, A. A., and Simoni, S. (2022). Renewable en-
Bharati, P., Narayan, K., and Ramachandra Rao, B. (2022). ergy, urbanization, fossil fuel consumption, and eco-
Green hydrogen production by water electrolysis: A nomic growth dilemma in Romania: Examining the
renewable energy perspective. Mat. Today: Proc., 67, short- and long-term impact. Energies, 15(19), 7180.
1310–1314. [Link]
5 Deep learning-based finger vein recognition and security:
A review
Manpreet Kaura, Amandeep Verma and Puneet Jai Kaur
Information Technology, UIET, Punjab University, Chandigarh, India
Abstract
The recognition system implies development that passes through the various stages. The finger vein recognition (FVR) is the
lead over the other biological modalities like finger print, face iris, etc., this paper reviews the worked done in the area of
FVR. The pre-processing is needed to enhance the images for better results. The feature extraction module provides the col-
lection of the best features in the finger vein images, which is used for template generation. The template-based schemes are
in fact the best and most appropriate for security purpose because it only preserved the scrambled information rather than
the original features of the human beings. According to this review the convolutional neural network (CNN)-based models
are best for FVR but still there have been some challenges faced by it so those would be improved in the experimental work
of this review.
manpreet.09bhagat@[Link]
a
Applied Data Science and Smart Systems 35
Table 5.1 Comparison of different biometric modalities
hackers because the template used by the scheme is of sparse because all the training samples are based
in the form of plain data (Qin and El-yacoubi, 2017). on dictionary matrix. Thus, need was felt to optimize
The author proposed the FVR-DLRP to secure revo- the selection of dictionary data to improve the system
cable and to do efficient finger vein template genera- performance (Fang et al., 2022). Finger vein recog-
tion. Most of the traditional finger vein recognition nition was done with the use of oval PDCNN, this
systems have a shading and misalignment of finger was the advanced version of the PDKs. It provided the
vein problem. Need to pay the much more effort best performance, the first ten layers were from the
and time for extracting the features from the images MobileNet and all other layers from the SqueezeNet
which is a complicated and complex process in deep network could be compressed the network to achieve
CNN (Y. Liu et al., 2018). To improve these problems the higher performance (Li et al., 2023).
the researchers had proposed a robust CNN model The CNN models have improved the recognition
which had the error rate of 0.396 it was collected on performance for FVR. But this advancement has
a good quality dataset (Hong, Lee, and Park, 2017). needed to improve the feature extraction module and
All the publicly available datasets for finger vein have defense against security attacks, and the CNN models
the small images collection. are not lightweights.
Although, the CNN model used for the finger vein,
achieved higher accuracy yet it face the problem Security of FVR system
related to the training process (W. Liu et al., 2017). In the present time, the security of each system is at
The CNN model is successfully applied for the fin- major risk of losing the information and digital assets
ger vein identification process (Simonyan, 2019). The which are protected using the several security mecha-
light weight convolutional neural network (CNN) nisms. Certain things are to be kept in mind before
model was proposed to improve the small training choosing the bio-information in any application to
dataset problem by using the similarity measure net- enhance the security. First of all, it should be clear
work (Qin and El-yacoubi, 2018). The dense net was that, all the phases of technologies which are to be
proposed for finger vein recognition (FVR). This was used in the application must be safe in terms of data
used it remove the noise in the images and the two privacy and protection. Although biometric technol-
or more images were combined for recognition sys- ogy yet it faces various security attacks is considered
tem and also used for feature extraction. The system to be a secure system of real-world market.
has huge computational cost (Song, Kim, and Park,
2019). Although the all-available methods were tested Template protection for finger vein
on publicly available datasets but still these systems The template protection is the process of generating
had failed in the practical uses. Depth based separate the precise or related information from the images by
CNN model was developed to overcome this prob- feature extraction process (Kumar, 2019). This is the
lem. The system is simple but it still has weakness of unique precise information associated with the differ-
recognizing the less defined features of the images ent human beings. The template scheme is divided into
(Tang et al., 2019). The CNN-CO was based on the two categories: (a) Bio-cryptosystem (Kaur, Kumar,
local descriptor for pre-training the ImageNet model. and Singh, 2023). This combines the best feature of
Practically this system was best out of all the con- both the worlds i.e., cryptographic keying methods
ventional systems (Y. U. Lu, 2019). Finger vein-based and biometric schemes and (b) cancellable templates
authentication model was proposed by developing the (Manisha, 2020). In this field, the biometric template
lightweight Siamese network. When images were col- of a person is distorted in such a manner that the
lected the feature information got lost. The GCNet original data is not available to the intruder but still
and multi-scale feature was used during the process to identity recognition can be performed. The template
solve the faced problem. The system has been tested generation in the CNN models for higher accuracy for
over three publicly available datasets but the need FVR can be performed (Yin, Zhang, and Liu, 2021).
was felt to improve the feature extraction module by The template generation is shown in Figure 5.2.
changing the width and cardinality of the CNN net- The security to the templates of finger vein to provide
work (Fang, Ma, and Li, 2023). The parameters for
training and testing the network for data happen to
be complex due to the complex hidden structure of
the neural network. Although the result was 99.98%
but we need a light weight model for FVR (Wang and
Shi, 2022). The double-weighted group sparse rep-
resentation classification was developed to solve the
FVR. It had the lower accuracy than the other mod-
els and also took a long time to solve the coefficient Figure 5.2 Template generation
Applied Data Science and Smart Systems 37
the password based key derivation function has been state of the art does address various issues related to
used to derive the key named FVR-DLRP. To provide the finger vein pattern recognition for artificial neural
the information still remains even though the pass- network. The question arises as who can suggest good
word is cracked. But this system has lower accuracy matching for prob and gallery images for recognition
in terms of FAR rate, which leads to FAR attack and process and has also been advised to prepare and
it even has lower GAR rate (Y. Liu et al., 2018). generate the light weight model for finger vein (Yin,
Deep CNN with hard mining finger verification Zhang, and Liu, 2021).
scheme was proposed which achieved better perfor-
mance than achieved through commercial finger vein Security attacks
verification systems. This method also accelerates the On the other side, the restricted system is responsible
complete training process. The huge template size for security attacks. The attackers generate fake bio-
requires enormous amount of storage space (Huang metric template and modify it at different levels, the
et al., 2017). The Gabor filter was used, built a fin- finger vein faces various security attacks related prob-
ger vein authentication system based on the light- lems, from time-to-time various security methods and
weight CNN and supervised discrete hashing so as their patches for breaches (Tome and Vanoni, 2014)
to improve the finger vein images. Despite all these are introduced to amend the system.
steps, this method has decreased the template size Transferable deep convolutional network was
and surges the performance of finger vein verification. proposed to handle the presentation attack. The
The connection between training time and recogni- researcher has modified the system by adding seven
tion outcome was not thoroughly measured (Xie and layers to the existing Alex-Net to overcome the over
Kumar, 2019). The fusion based system was devel- fitting problem. The artifacts for the finger vein were
oped for fingerprint and the finger vein biological generated by using two different printers. The modi-
datasets. The feature level fusion was applied to this fied system is able to handle the PAD for finger vein.
system. The system has the higher matching perfor- The transferable deep learning neural network for the
mance and security of the data (Yang et al., 2018). finger vein has still to face video presentation attack
The BDD-based FVR system was based on the deep (Raghavendra et al., 2017). Another mechanism for
CNN. The system was combined with ML-ELM to security of the data template was developed but it is
form a FVR system that provides the protected pri- still facing security breaches during different process
vacy, and also provides the security to the template. like storing the template and matching the templates
In case of tempering by the intruders the template for finger vein. The FVR with the template-based pro-
performs the undoing operation independently. Then tection is able to handle the presentation attack but
the new template version is generated by user specific still has the problem related to the adversarial attack
keys (Yang et al., 2019). The weighted least squares (Ren et al., 2021). The survey provided by the Yimin
regression has been used to improve the template gen- Yin et al. (Yin, Zhang, and Liu, 2021), had suggested
eration. It minimizes the verification errors but this and elaborated all the security breaches in finger vein
was based on an assumption, but the template has CNN methods. The system has developed to handle
the very little distance in the intra-class for the same the impersonation attack and check the system for
image data (Qin, 2019). The cancelable biometric- authentication with minimum enrolment time. The
based scheme for CIRF, and proposed a low-rank MC-CLAHE method is used to process the images
approximation-based cancelable indexing scheme of finger vein prior to the CNN training process but
which was based on CIRF was introduced to solve the the system is only providing security to the database
problem of excessive computational overhead. Low- (Safie, Zarina, and Khalid, 2023). Various templates-
rank approximation of biometric images was used based CNN models are developed. It has been pro-
to speed up the calculation of CIRF and also used vided the moderate defense against security attacks.
the minimum spanning tree representation for low- But the security problem still remains exist in the
rank matrices in the Fourier domain. The researcher template based FVR systems. Table 5.3 has shown the
proved the reliability of the projected method in pro- most recently articles related to the CNN, Template
tecting related biological information (Murakami et and security of FVR.
al., 2019). The template protection scheme has been
proposed to align the images. The IoM hash is used to IV. Datasets
realize the required privacy and security for the FVR
(Kirchgasser et al., 2020). The security is provided Various datasets (Y. Lu et al., 2013) are always freely
to the template to solve the issue of the presentation available to the researchers to train and test the pre-
attack by CNN model but templates still faces adver- trained models for the CNN. The models learn the
sarial sample attack and does not have light weight features from the images in the dataset to train the sys-
feature extraction module (Ren et al., 2021). The tem, then the system based on this training recognizes
38 Deep learning-based finger vein recognition and security: A review
Table 5.3 Recent articles related to the CNN and template security of FVR
Deep CNN DS1, DS2 & DS3 EER - DS1=0.42%, Improvement IEEE International (Huang et
DS2= 1.41% & DS3= of vein pattern Conference al., 2017)
2.14% matching on Identity,
Security and
Behavior Analysis
(IEEE-2017)
Deep learning FV_NET64 GAR=91.2% Enhancement the Soft Computing (Y. Liu et
& random FAR=0.3% revocability of the (Springer-2018) al., 2018)
projection template
EP-DFT FVC2002 DB2, EER=0.45% Enhancement of Pattern Recognition (Yang et al.,
FVC2004 DB2, non-revocability of (Elesvier-2018) 2018)
FV-HMTD templates
CNN and Two session EER=0.0887 Reduced the Pattern (Xie and
supervised databases template size Recognition Letters Kumar,
discrete hashing (Elesvier-2019) 2019)
BDD-ML-ELM SDUMLA, CIR=93.09%, 98.70%, New non invertible IEEE Transactions (Yang et al.,
MMCBNU_6000 98.61%, respectively templates for on Industrial 2019)
& UTFVP (datasets) finger vein Informatics
(IEEE-2019)
Weighted HKPU & FV-USM EER=1.28, 1.43, Improvement in MDPI (Qin, 2019)
least square respectively (datasets) enrolment template (Information-2019)
regression for finger vein
Correlation- N Genuine template Fast and secure Pattern (Murakami
invariant identification=164.7 biometric Recognition Letters et al., 2019)
random identification 7 (Elesvier-2019)
filtering No leakage of
information of the
template
40 Deep learning-based finger vein recognition and security: A review
Self-attention FV-USM, Recall is best over Need to improve Infrared Physic (Fang, Ma,
mechanism MMCBNU_6000, SDUMLA= 0.9944 the feature and Technology and Li,
(SAC) Siamese SDUMLA-HMT F1 Score is best over extraction module (Elsevier Dec-2020) 2023)
network FV_USM=0.9925 by changing
EER is best over the width and
MMCBNU_6000= cardinality of the
0.0012 network models
RSA for SDMLA, Scheme A is best over The feature Knowledge Based (Ren et al.,
template MMCBNU_6000, HKPU=99.03% extraction method System (Elsevier 2021)
protection HKPU, FV-USM and is not lightweight May-2021)
using CNN Scheme B is best over and the system is
FV_USM=99.18% not able to handle
the adversarial
sample attack and
presentation attack
Survey on SDUMLA-HMT, Performance summary Need to generate the Computer Vision (Yin,
ANN for FV-USM, HKPU, of all CNN methods light weight models, and Pattern Zhang, and
finger vein MMCBNU_6000, for finger vein solve the problem Recognition Liu, 2021)
UTFVP, THU- of mismatching (Springer
FVFDT, SCUT, of gallery and Aug-2022)
IDIAP prob image,
dynamic finger vein
extraction
Double PolyU [], FV-USM, The system has best The sparse International (Fang et al.,
weighted SDUML-HMT performance over coefficient takes the Journal of Machine 2022)
group sparse FV_USM=97.0-95.88- long time, so that Learning and
representation 88.72% over three need to improve it Cybernetics
classification different variants of and improve the (Springer
models accuracy of the May-2022)
system
Multimodal CASIA-WebFace, 99.98% Model should be Sensors (MDPI (Wang and
approach SDUMLA-FV, lightweight Aug-2022) Shi, 2022)
based on CNN FV-USM
(RESNET,
AlexNet,
VGG-19)
MC-CLAHE FV-USM AUC=0.78–0.91 Security only International (Safie,
(CNN- provided to the Journal of Online Zarina,
AlexNet) database and Biomedical and Khalid,
Engineering (iJOE 2023)
2023)
N=Data not available.
Abstract
Fabrication of devices in deca nanometer regime suffers from several limitations as the devices are being scaled so that the
speed and transistor density can be increased. This has led to a series of innovative techniques by the industry as well as aca-
demia. Depletion regions formed in association with the p-n junctions is one of the restrictive factors in scaling short channel
devices in case of junction-based (JB) metal-oxide-semiconductor field-effect transistors (MOSFETs). This has led to several
short channel effects (SCEs). Recently, novel MOSFET structures have been developed that are devoid of p-n junctions and
have also been successfully fabricated. These devices are named “junctionless transistors (JLTs)”. MOSFETs employing
gate-all-around (GAA) architecture have been reported as an ultimate structure in silicon integrated circuits (ICs). In this
paper, we have developed an analytical drain current model for short channel GAA JLT, including source (S)/drain (D) series
resistance, which is also one of the important parameters when devices with short channel are fabricated. We have obtained
the potential distribution profile using Poisson’s equation. It was then used for obtaining the model for drain current. The
validation of the model has been obtained with both the simulation as well as experimental results. We have further analyzed
the effect of S/D resistance on the drain current for different device parameters.
a
[Link]@[Link]
44 Development of an analytical model of drain current for junctionless GAA MOSFET
II. Theoretical details voltage (VFB) of the device, a complete neutral chan-
nel is created and we reach flat band condition. The
GAA structures offer superior short-channel charac-
conduction and valence bands become flat and now
teristics owing to the excellent control the gate offers
we can say that the device is turned ON. An accumu-
over the channel in such structures. Due to the absence
lation layer is created at the surface on further increas-
of junctions in JLTs, there is no requirement of doping
ing the gate voltage and the negative charge carriers
concentration gradient and hence the problems asso-
get accumulated resulting in the flow of surface cur-
ciated with the junctions are eliminated. Such devices
rent along with the bulk current. Figure 6.2 depicts
are also reported to deliver improved driving current
the energy-band diagram of GAA JLT in different
and sub-threshold properties when combined with
regions of operation along with the device schematic.
GAA architecture. Figure 6.1 depicts the cross-section
For an n-type semiconductor, in the cylindrical
of such JL GAA MOSFET.
coordinate, the Poisson’s equation can be written as
Device physics
JLTs are characterized as devices that are strongly and
(1)
evenly doped throughout. This means the type of the
dopants and their concentration is same all over the
junction. For n-type devices, p+ polysilicon is used as
where, φ signifies the potential,
the gate material and n+ polysilicon is used for p-type
r represents the radial direction,
devices. This results in a difference of approximately
V is the applied voltage,
1 eV in the work function which causes the channel
Nd represents the concentration of dopant throughout
to deplete. To bring the channel out of depletion, gate
the source, drain and channel,
bias must be applied. The working principle of GAA
VT is thermal voltage, and
JLT is as follows.
εSi is the permittivity of Si.
When no gate voltage (VG) is applied, the channel
Neglecting depletion charge density and considering
is fully depleted and a negligible amount of current
only the mobile carrier’s density, the above equation
flows through the region between the S and D. In such
can be simplified (Trevisoli et al., 2012) as
a situation, the transistor is said to be in OFF con-
dition. This is the sub-threshold region of operation.
The valence band is completely filled while the con- (2)
duction band is empty. When the voltage applied at
the gate equals the device’s threshold voltage (VTH),
bulk current starts flowing along a thin neutral path,
which is a non-depleted region formed near the center
of the channel. The path gets widened on increasing
the gate voltage. This increases the current flowing
through it. This has been reflected in the energy band
diagram where it can be seen that the concentration of
positive charges in the valence band is reduced. When
the applied voltage at the gate is same as the flat-band
(3)
(4)
(5)
(6)
(7)
(8)
(9)
46 Development of an analytical model of drain current for junctionless GAA MOSFET
before going into final fabrication. This may save parameters of nanoscale strained silicon MOSFET-
time as well as resources. We have not taken quantum based CMOS inverters. Microelec. J., 55, 8–18.
mechanical effects into account which becomes sig- Subindu, K., Kumari, A., and Das, M. K. (2017). Model-
nificant in ultra scaled devices. ing gate-all-around Si/SiGe MOSFETs and circuits for
digital applications. J. Comput. Elec., 16, 47–60.
Amrita, K., Saini, A., Kumar, A., Kumar, V., and Kumar,
References M. (2023). Recent developments and challenges in
strained junctionless MOSFETs: A review. 2023 Int.
Rishu, C. and Yirak, M. G. (2023). Sensitivity investigation
Conf. Comput. Intel. Sustain. Engg. Sol. (CISES),
of junctionless gate-all-around silicon nanowire field-
118–122.
effect transistor-based hydrogen gas sensor. Silicon,
Haijun, L., Zhang, L., Zhu, Y., Lin, X., Yang, S., He, J., and
15(1), 609–621.
Chan, M. (2012). A junctionless nanowire transistor
Te-Kuang, C. (2012). A new quasi-2-D threshold voltage
with a dual-material gate. IEEE Trans. Elec. Dev.,
model for short-channel junctionless cylindrical sur-
59(7), 1829–1836.
rounding gate (JLCSG) MOSFETs. IEEE Trans. Elec.
Dong-Il, M., Choi, S.-J., Duarte, J. P., and Choi, Y.-K.
Dev., 59(11), 3127–3129.
(2013). Investigation of silicon nanowire gate-all-
Sung-Jin, C., Moon, D., Kim, S., Ahn, J.-H., Lee, J.-S., Kim,
around junctionless transistors built on a bulk sub-
J.-Y., and Choi, Y.-K. (2011). Nonvolatile memory by
strate. IEEE Trans. Elec. Dev., 60(4), 1355–1360.
all-around-gate junctionless transistor composed of
Sehra, S. S., Singh, J., Rai, H. S., and Anand, S. S. (2020).
silicon nanowire on bulk substrate. IEEE Elec. Dev.
Extending processing toolbox for assessing the logi-
Lett., 32(5), 602–604.
cal consistency of OpenStreetMap data. Trans. GIS,
Jean-Pierre, C. (2007). Multi-gate SOI MOSFETs. Micro-
24(1), 44–71. [Link]
elec. Engg., 84(9–10), 2071–2076.
Pratikhya, R. and Nanda, U. (2022). A charge-based analyt-
Jean-Pierre, C., Lee, C.-W., Afzalian, A., Akhavan, N. D.,
ical model for gate all around junction-less field effect
Yan, R., Ferain, I., Razavi, P. et al. (2010). Nanow-
transistor including interface traps. ECS J. Solid State
ire transistors without junctions. Nat. Nanotechnol.,
Sci. Technol., 11(5), 051006.
5(3), 225–229.
Pushpapraj, S., Singh, N., Miao, J., Park, W.-T., and Kwong,
Duarte, J. P., Choi, S.-J., Moon, D., and Choi, Y.-K. (2011).
D.-L. (2011). Gate-all-around junctionless nanowire
A nonpiecewise model for long-channel junctionless
MOSFET with improved low-frequency noise behav-
cylindrical nanowire FETs. IEEE Elec. Dev. Lett.,
ior. IEEE Elec. Dev. Lett., 32(12), 1752–1754.
33(2), 155–157.
Billel, S., Rahi, S. B., and Labiod, S. (2022). Analytical com-
Philippe, G., Teramoto, A., and Ohmi, T. (2010). Modelling
pact model of nanowire junctionless gate-all-around
of the hole mobility in p-channel MOS transistors fab-
MOSFET implemented in verilog-A for circuit simula-
ricated on (1 1 0) oriented silicon wafers. Solid-State
tion. Silicon, 14(16), 10967–10976.
Elec., 54(4), 420–426.
Billel, S., Nafa, F., Upadhyay, A. K., Labiod, S., Rahi, S. B.,
Guangxi, H., Xiang, P., Ding, Z., Liu, R., Wang, L., and
Benlatreche, M. S., Akroum, H., Lakhdara, M., and
Tang, T.-A. (2014). Analytical models for electric po-
Yadav, R. (2023). Compact modeling of junction-
tential, threshold voltage, and subthreshold swing of
less gate-all-around MOSFET for circuit simulation:
junctionless surrounding-gate transistors. IEEE Trans.
Scope and challenges. Device Circuit Co-Design Is-
Elec. Dev., 61(3), 688–695.
sues in FETs, 57–78. CRC Press.
Chunsheng, J., Liang, R., Wang, J., and Xu, J. (2014). Ana-
Trevisoli, R. D., Doria, R. T., de Souza, M., Das, S., Ferain,
lytical short-channel behavior models of junctionless
I., and Pavanello, M. A. (2012). Surface-potential-
cylindrical surrounding-gate MOSFETs. 2014 Int.
based drain current analytical model for triple-gate
Symp. Next-Gen. Elec. (ISNE), 1–2. IEEE.
junctionless nanowire transistors. IEEE Trans. Elec.
David, J., Iniguez, B., Sune, J., Marsal, L. F., Pallares, J.,
Dev., 59(12), 3510–3518.
Roig, J., and Flores, D. (2004). Continuous analytic
Tsormpatzoglou, A., Tassis, D. H., Dimitriadis, C. A.,
IV model for surrounding-gate MOSFETs. IEEE Elec.
Ghibaudo, G., Pananakakis, G., and Clerc, R. (2009).
Dev. Lett., 25(8), 571–573.
A compact drain current model of short-channel cy-
Ye-Ram, K., Lee, S.-H., Sohn, C.-W., Choi, D.-Y., Sagong,
lindrical gate-all-around MOSFETs. Semicond. Sci.
H.-C., Kim, S., Jeong, E.-Y. et al. (2013). Simple S/D
Technol., 24(7), 075017.
series resistance extraction method optimized for
Juncheng, W., Du, G., Wei, K., Zhao, K., Zeng, L., Zhang,
nanowire FETs. IEEE Elec. Dev. Lett., 34(7), 828–
X., and Liu, X. (2014). Mixed-mode analysis of differ-
830.
ent mode silicon nanowire transistors-based inverter.
Alok, K., Gupta, T. K., Shrivastava, B. P., and Gupta, A.
IEEE Trans. Nanotechnol., 13(2), 362–367.
(2023). Impact of temperature variation on noise
Yu, Y. S. (2014). A unified analytical current model for
parameters and HCI degradation of recessed source/
N-and P-type accumulation-mode (junctionless) sur-
drain junctionless gate all around MOSFETs. Micro-
rounding-gate nanowire FETs. IEEE Trans. Elec. Dev.,
elec. J., 134, 105720.
61(8), 3007–3010.
Subindu, K., Kumari, A., and Das, M. K. (2016). Develop-
ment of a simulator for analyzing some performance
7 Crop recommendation using machine learning
Paramveer Kaura and Brahmaleen Kaur Sidhu
Department of Computer Science and Engineering, Punjabi University, Patiala, Punjab, India
Abstract
Agriculture serves as the cornerstone of India’s economic expansion, constituting the primary income source for a significant
proportion of its populace, encompassing both those directly engaged in agricultural activities and those who depend on it
indirectly for their livelihoods. Therefore, it is essential for farmers to make the correct option possible when cultivating any
crop so that the farmer can make maximum profit from the agriculture field. To make the agriculture sector profitable, one
of the technologies that may be used in this age of rapid technological improvement is known as machine learning (ML). In
this research paper, various ML algorithms, such as logistic regression (LR), decision trees (DT), LightGBM, and random
forest (RF), have been utilized to analyze a dataset. The primary objective is to predict the most suitable crop based on soil
attributes such as (nitrogen, phosphorous, potassium) NPK content, humidity, temperature, soil pH level, and rainfall. Out of
this random forest and LightGBM comes with great accuracy whereas decision tree and logistic regression have less accuracy.
In addition, ML algorithms will likely find applications in a variety of agricultural subfields in the near future, including the
diagnosis of plant diseases, the selection of soil types, and the forecasting of retail pricing.
Keywords: Machine learning, crop recommendation, decision tree, random forest, logistic regression, LightGBM
a
paramveer1067@[Link]
50 Crop recommendation using machine learning
this four, ML models are deployed to make accurate farmers’ crop management issues like crop selection,
crop recommendations. yield, and profit. Researchers employ decision trees,
Naive Byes, SVM, LR, RF, and Xgboost. Pradeepa
II. Related work Bandara et al. developed a crop recommendation sys-
tem for Sri Lanka (Bandara et al., 2020). The study
Pudumalar et al. (2016) addressed precision agricul- provides a theoretical as well as a conceptual plat-
ture. This study proposes an ensemble model with form for a recommendation system using Arduino
majority voting technique utilizing RT, Naïve Bayes, microcontrollers, ML approaches such as Naive
CHAID, and K-nearest neighbor (KNN) as learn- Bayes (multi-nomial) and SVM and unsupervised ML
ers to effectively and correctly suggest a crop for algorithms that are K-Means Clustering and Natural
site-specific parameters. The study (Kanaga Suba Language Processing (NLP) (sentiment analysis).
Raja et al., 2017) analyses historical data to predict Avinash Kumar et al. (2019) addressed crop selection
a farmer’s crop output and price. Sliding window and disease issues. SVMs classification model, deci-
non-linear regression predicts agricultural output sion tree model, and logistic regression model were
depending on rainfall, temperature, market prices, used to create this recommendation system.
land area, and crop yield. Zeel Doshi et al. (2018)
developed a soil dataset-based crop recommendation
III. Objectives
system for only four crops. The ensemble model uses
random forest (RF), Naive Bayes, and linear support The aim of the proposed work is to implement ML
vector machines base learners. The majority voting algorithms for developing crop recommendation sys-
technique is employed in the combination approach tem and it is based on chemical properties of soil and
because it is the most accurate. The author uses Big weather condition and to evaluate the proposed system.
Data analytics and ML to create an AgroConsultant,
an system that assists Indian farmers choose the best IV. Background techniques
crop based on sowing season, farm location, soil
properties, and environmental factors like tempera- A. Logistic regression
ture and rainfall (Doshi et al., 2018). Rainfall predic- It is a ML algorithm primarily applied to classifica-
tor, another approach created by academics, predicts tion problems, operates on the foundation of predic-
annual precipitation. The system uses DT, KNN, RF, tive analysis rooted in probability (Rymarczyk et al.,
and neural networks. 2019). Notably, the LR model adopts a more intricate
Shilpa Mangesh Pande et al. (2021) provide farm- cost function than linear regression. This cost func-
ers a simple yield projection tool. Farmers utilize a tion, often referred to as the “Sigmoid function” or
smartphone app to connect to the internet. GPS “logistic function,” replaces the linear function. Due
locates users. Enter location and soil type. ML sys- to the fundamental premise of logistic regression, the
tems can identify the most profitable crops and pre- output range of the cost function is inherently con-
dict agricultural yields for user-selected crops. SVM, fined to the interval [0, 1]. This constraint stems from
RF, MLR, ANN, and KNN are used to predict agri- the nature of logistic regression, where it inherently
cultural production. The research (A et al., 2021) models the probability of an event occurring, ren-
suggests a way to assist farmers pick crops by con- dering linear functions unsuitable for capturing its
sidering planting time, soil qualities including type of nuances and characteristics.
soil, pH value, and nutrient content, meteorological
aspects like rainfall, temperature, and state location. B. Decision tree
The suggested system has been developed using linear When it comes to representing models for use in
regression as well as neural network. Another study data classification, decision trees are among the most
(Gosai et al., 2021) forecasts the best crop based on popular approaches (Jijo and Abdulazeez, 2021). DT
N, P, K, pH of soil, humidity, temperature, and rain- stand as versatile assets in numerous domains, span-
fall. Decision trees, SVM, Nave Bayes, Support vec- ning machine learning, image processing, and pattern
tor machine, LR, RF, and XGBoost were utilized to recognition. Their core role revolves around the task
develop suggested system, and the maximum accu- of classification and, as a result, they find wide appli-
racy was of XGBoost. Distribution analysis, majority cation as classifiers within the field of data mining.
voting, correlation analysis and ensembling are used These decision trees are architecturally composed of
to create 22 crop recommendations (Kulkarni et al., interconnected nodes and branches, each node rep-
2018). A three-level technique solves crop recommen- resenting a collection of attributes within discrete
dations. Chhikara et al. (2022) propose a ML-based classification categories. Each branch within the tree
crop recommender system that can accurately fore- signifies a potential value associated with the respec-
cast the yield of 22 different crop types, addressing tive node. Decision trees earn considerable favor for
Applied Data Science and Smart Systems 51
C. Random forest
In the field of ML, random forest, a supervised learn-
ing technique, has demonstrated considerable success.
It’s versatile and can handle various tasks like classifi-
cation and prediction. What sets random forest apart
is that it operates like a team of decision trees collabo-
rating to solve problems. Rather than rely on a one
decision tree, it combines the results of many trees,
each trained on different parts of the data. This coop-
erative approach enhances accuracy. Essentially, the
more trees in this “forest,” the better the performance,
and it’s less prone to errors (Dabiri et al., 2022; Gera
et al., 2021).
D. LightGBM
Tree-based learning algorithms are at the foundation
of LightGBM, a gradient boosting framework (Tang
et al., 2020). It has the many benefits because of its
decentralized and efficient design such as increased
efficiency and accelerated training time. It uses less
memory and allows for multi-GPU and distributed
learning. LightGBM has ability to process large Figure 7.1 Flow chart of proposed methodology
amount of data as well as gives enhanced precision.
Abstract
Artificial Intelligence (AI) and sustainability are two sides of same coin. AI is a reliable ally in the fight for sustainability,
leading us to a brighter future. AI illuminates renewable energy, resource management, and eco-friendly decision-making by
analyzing large datasets. However, the energy usage and carbon footprint of AI models and AI sustainability are increasingly
under review. This research paper examines the environmental implications of AI models, focusing on ChatGPT, and empha-
sizes the necessity for sustainable AI development. Recent studies show that AI model creation and use significantly impact
the global carbon footprint due to energy, water, and carbon emissions. With its massive computational needs, ChatGPT con-
tributes to environmental issues. To tackle this dilemma, sustainable AI development must be promoted. Model compression,
quantization, and knowledge distillation improve AI energy efficiency. The use of renewable energy and the establishment
and enforcement of AI model energy efficiency requirements are equally crucial. ChatGPT and comparable models can be
environmentally friendly by using sustainable AI development methods. In this line, the objective of the present study is to
analyze the impact of the use of AI tools, specifically ChatGPT, on sustainability and environmental protection by analyzing
existing reports and studies on the environmental impact of artificial intelligence models.
Academicians, developers, politicians, institutions and organizations must work together to create rules and frameworks
for energy-efficient AI algorithms, renewable energy use, and responsible deployment. This study article concludes that AI
models’ energy usage and carbon footprint must be understood and reduced. By promoting sustainable practices, the AI
community may encourage a more environmentally sensitive and responsible approach to AI development, leading to a
greener future that meets global sustainability goals.
Keywords: Artificial intelligence, AI-language models, ChatGPT, environmental impact, carbon footprint, water footprint,
greenhouse gas emissions, energy consumption, sustainability, mitigation strategies
neha_seth01@[Link]
a
Applied Data Science and Smart Systems 55
models by addressing such challenges. Suggestions for significantly impacts business sustainability. Ray
the adoption of sustainable practices and the usage of (2023) asked ChatGPT how it will play a significant
renewable energy sources in all fields are part of the role in agricultural science and technology in future
study. This article addresses the environmental effect and got responses which may lead to the sustainable
of ChatGPT and offers suggestions for sustainable AI development of the farming sector.
development, but it also has certain restrictions and After refereeing a number of research papers and
opens up avenues for future exploration. articles published on ChatGPT and its application in
The study highlights the need to take into account various fields, it was found that individually many
the ecological implications of AI systems and the papers talk about the application of ChatGPT in vari-
demand for sustainable practices in creating and ous areas for sustainable development and extend-
implementing them. The last part includes the conclu- ing similar work, in this study, authors are trying to
sion and future prospects of the study. analyze the impact of the use of AI tools, specifically
ChatGPT, on sustainability and environmental pro-
II. Review of literature tection by analyzing existing reports and studies on
the environmental impact of artificial intelligence
In the contemporaneous literature available in the models.
field of information technology, numerous studies are
available that provide information about the chatbots,
III. ChatGPT and potential areas of concern
language models and IT platforms. This section pres-
ents the evolvement of research based on ChatGPT Kain (2023) stated that ChatGPT is an advanced
and its relation to sustainable development. language model developed by OpenAI and released
Zhu et al. (2023) raised concerns about environ- in November 2022. The acronym “ChatGPT” com-
mental issues due to the introduction of another natu- bines the terms “Chat”, which refers to the chatbot
ral language processing model, ChatGPT. They have functionality of the framework, and “GPT” stands for
used ten real-world examples to study the impact of “Generative Pre-trained Transformer” and it is a type
ChatGPT and its impact on the environment. Another of Large Language Model (LLM). ChatGPT is based
study by Khowaja et al. (2023) focused on an aspect on the core GPT models from OpenAI, GPT-3.5 and
of large language models which were ignored, includ- GPT-4, which provide conversational interaction
ing sustainability, privacy, digital divide and ethics (Lund et al., 2023).
(SPADE) and based on primary data and visualization, To generate intelligent and captivating text-based
they suggested that not only ChatGPT other models replies to user input, the latest AI conversation tool
should also undergo SPADE analysis. George (2023) takes advantage of the most recent advancements in
raised the issue of water consumption by ChatGPT. It machine learning as well as natural language process-
was found that the water consumption by AI models ing (NLP) (Bhaskar, 2022).
is relatively less than in other industries but it is still OpenAI launched in November 2022, ChatGPT
a matter of concern, and it should be further reduced revolutionized how people interact globally by pro-
by taking appropriate measures like improving energy ducing replies to common writing jobs in seconds.
efficiency, utilizing renewable energy sources, opti- Despite the fact that ChatGPT’s “outputs may be
mizing algorithms and implementing strategies to inaccurate, untruthful, and otherwise misleading at
conserve water. times”, as stated in its FAQs, the model’s speed and
Biswas (2023) integrated with ChatGPT to get adaptability make it widely applicable for simple
responses on the effect of ChatGPT on global warm- writing tasks, including cover letters, and many more
ing and analyzed the replies received. Biswas con- uncountable things.
cluded that ChatGPT can be used in various ways to ChatGPT does not directly impact the environ-
aid climate research, including model “parameteriza- ment because it is an AI language model. However,
tion, data analysis and interpretation, scenario gen- the infrastructure and data centers needed to support
eration, and model evaluation”. Sohail et al. (2023) ChatGPT, as well as the technology that enables it, may
reviewed 100 Scopus papers on ChatGPT and found have an impact on the environment. Such as training
that ChatGPT has applications in various fields like and operating complex language models like GPT-3
healthcare, marketing and financial services, software require a substantial amount of computer power,
engineering, academic and scientific writing, research which is why ChatGPT uses a lot of energy. If the
and education, environmental science, and natural energy required to power data centers and computer
language processing and its potential to address real- infrastructure comes from non-renewable sources, it
world problems. Vrontis et al. (2023) analyzed the may cause carbon emissions and environmental dam-
role of ChatGPT and skilled employees in business age (Teubner, 2023). Also, the energy needed to run
sustainability and found that leadership motivation AI models results in the emission of carbon dioxide
56 Environment and sustainability development: A ChatGPT perspective
and other greenhouse gases, which fuel global warm- ChatGPT’s 1 million users sent one request each day,
ing. The carbon footprint of ChatGPT and other AI ChatGPT would get the same number of requests per
systems depends on factors such as the energy source, day as BLOOM did at that time. At least based on the
cooling requirements, and hardware efficiency (An et volume of discussion about ChatGPT in traditional
al., 2023; Khattar et al., 2020). and social media platforms, ChatGPT handles far
ChatGPT needs a lot of processing and storage more daily requests than similar services. Although
power, which is often provided by big data centers. there is a great deal of ambiguity around this estimate
Massive amounts of water are used to cool these data since it is founded on some dubious presumptions,
centers, which leads to a water footprint. Apart from compared to in-depth analyses of BLOOM’s carbon
that, they also need various other materials like metals footprint, a comparable linguistic model, it seems
and minerals to build and maintain them (Qin, 2023). plausible.
As AI technology develops quickly and becomes obso- Figure 8.1 ([Link]) shows the energy con-
lete, it may cause an increase in e-waste as outmoded sumed by AI models during training is significant, with
gear is discarded. E-waste poses environmental risks both GPT-3, the first version of the current edition of
due to improper disposal, as it often contains hazard- OpenAI’s popular ChatGPT, and Gopher requiring
ous and toxic elements (Khan, 2023). well over a thousand-megawatt hours of electricity.
Because this is solely for the training model, the over-
IV. ChatGPT’s: Creating carbon footprint all energy consumption of GPT-3 and other LLMs is
expected to be substantially greater. GPT-3, the great-
The phrase “carbon footprint” denotes the total est energy user, consumed nearly the equivalent of
quantity of “carbon dioxide (CO2)” pollutants gener- 200 Germans in 2022. While not enormous, it repre-
ated by a particular person or a company (such as a sents a significant usage of energy.
nation, business, building, etc.). Both immediate emis- While it is undeniable that training LLMs requires
sions from the energy generation that drives consumer a significant amount of energy, the energy sav-
products and services and further emissions from the ings are anticipated to be significant. Any AI model
burning of fossil fuels for industry, transportation, that improves operations by a fraction of a second
and heating make up this total. Additionally, meth- might save hours of shipping, liters of gasoline, or
ane, nitrous oxide, and chlorofluorocarbons (CFCs) hundreds of computations. Each consumes energy,
emissions are frequently taken into consideration and the total amount of energy saved by an LLM
when discussing a carbon footprint idea (An et al., may considerably outpace its energy cost. Mobile
2023; Euronews, 2023). phone carriers are an excellent example, with one-
Kain (2023) stated that the carbon footprint of third expecting AI to lower power usage by 10 per
creating ChatGPT isn’t public information, but if cent to 15 per cent. Given how much of the world
understood correctly, it is based on a GPT-3 variation. relies on mobile phones, this would be a significant
Estimates show that training GPT-3 consumed 1,287 energy saver. The CO2 emissions from training LLMs
MWh and generated 552 tons of CO2. are also significant, with GPT-3 emitting over 500
Mclean’s (2023) research paper stated that it is tons of CO2. This, too, might be drastically altered
most certainly considerably greater than that of GPT- depending on the sorts of energy generation that
3. The energy expenditures would increase if it had to cause the emissions. Most data center operators,
be rebuilt frequently in order to refresh its knowledge. for example, would want to have nuclear energy, a
The amount of carbon dioxide that ChatGPT is esti-
mated to produce annually is 8.4 tons, which is more
than twice as much as the annual emissions of a single
person. i.e., 4 tons.
ChatGPT’s daily emissions of 23.04 kg of CO2 per
day would add up to 414.72 kg CO2 during the course
of 18 days and on the contrary, Big Science Large
Open-science Open-access Multilingual Language
Model (BLOOM) (Luccioni, 2022) released 360
kg of CO2 during the course of 18 days. The differ-
ence between the two emission estimates can be due
to many things, including the varying carbon inten-
sities of the electricity (Jiafu, 2023) generated by Figure 8.1 Power usage for training large language
BLOOM and ChatGPT. It’s also crucial to remember models (LLMs) based on AI in 2023 (in megawatt
that BLOOM handled 230,768 requests in total over hours)
18 days, or 12,820 on average every day. If 1.2% of Source: Statista
Applied Data Science and Smart Systems 57
notably low-emission energy generator, play a big well over a thousand-megawatt hours of electricity.
role. Because this is solely for the training model, the over-
Figure 8.2 ([Link]) shows how the develop- all energy consumption of GPT-3 and other large lan-
ing world is making the most aggressive efforts to guage models (LLMs) is expected to be substantially
reduce emissions from AI models used in organiza- greater.
tions. This covers enormous regions in India, Africa, Figure 8.4 ([Link]) shows CO2 emissions
Latin America, and the Middle East, including a siz- from AI are significant compared to the average
able chunk of the planet and most of its inhabitants. human emission in 2022. GPT-3 training, not includ-
The proportion of organizations tackling emissions in ing the model that is now operating, produced more
those regions is approximately double that of North than a hundred individuals in a year. Training Gopher
America. North American and European organiza- was the equivalent of seventeen Americans’ emissions.
tions are taking fewer steps, which might be due to While the figure may appear considerable, it must be
the fact that the energy utilized to power this technol- seen from the perspective of potential emission reduc-
ogy on those continents is often greener than in the tions through more efficient business strategies.
developing world. It is observed from Figure 8.5 ([Link]) that
Figure 8.3 ([Link]) shows the energy con- the power consumption while training AI-based LLM
sumed by AI models during training is significant, with
both GPT-3, the first version of the current edition of
OpenAI’s popular ChatGPT, and Gopher requiring
Figure 8.3 Emissions when training AI-based large lan- Figure 8.5 Power consumption when training AI based
guage models (LLMs) in 2022 (in CO2 equivalent tons) large language models (LLMs)
58 Environment and sustainability development: A ChatGPT perspective
nevertheless, providers have a number of options for Figure 8.7 ([Link]) shows that when AI tools
reducing their digital footprint. were employed, organizations in 2022 were primarily
Table 8.1 represents the steps that are effective in concerned with reducing their physical influence on
reducing the environmental impact of Chat GPT: the environment. This is most likely owing to such
These steps are effective in reducing the environ- enhanced efficiency simply translating to improved
mental impact of ChatGPT. However, their effective- corporate growth and expenses. Organizations priori-
ness depends on the specific circumstances of the data tize ethical product sourcing since it may sometimes
center where ChatGPT is located. Some solutions result in direct cost increases to manufacturing and
may be more feasible than others based on the loca- supply lines.
tion of the data center, the type of hardware used, and
other factors.
For example, optimizing the location of a data
center may not be feasible in all cases. Data centers
may be located in areas with limited access to cooler
climates or water sources. In these cases, alternative
cooling methods or improving energy efficiency may
be more feasible solutions (Mclean, 2023).
Choose required information The cost of training a model may be greatly decreased by utilizing just the necessary
data or by successfully adapting current models for a new purpose, making AI more
viable
Invest in green energy In order to decrease CO2 emissions, efforts are needed to increase the use of renewable
energy sources in data centers. Therefore, it is advisable and essential to rely on cloud
service providers to make sure that electricity is delivered from renewable energy
Reduce unnecessary Unnecessary computations should be reduced in order to lower the overall workload of
computations ChatGPT. By improving the model’s data processing methods and algorithms, this can
be accomplished. Less energy and water will be needed to power ChatGPT by reducing
the amount of computing that is not required
Monitoring and analyzing It’s crucial to routinely track and evaluate ChatGPT’s water usage. Data center
water consumption operators may use this to streamline their processes and find places where water usage
can be decreased
Advocate for greater The creation and maintenance of measurements and standards for assessing the energy
transparency efficiency of creating and implementing ML models is one approach to resolving this
problem (Henderson et al., 2020)
Optimize data center location Locate and promote areas with the potential to host greater amounts of renewable
energy and data centers with reduced carbon footprints
Prolonging life of AI models Increasing the longevity of AI technology and infrastructure through upkeep,
maintenance, and updates can cut down on the production of electronic waste. To
minimize environmental impact, it is also crucial to recycle and properly dispose of old
AI technology
Encourage environmental Environmental awareness is critical because it will help promote responsible AI
awareness industry practices that will open the door for “greener” AI. One tactic for doing this is
to emphasize the limitations of language models and to lessen the excitement around
novel, eye-catching AI systems like ChatGPT. We may actively support new lines of
inquiry that do not simply rely on creating more complicated ones (Zhu, 2023)
60 Environment and sustainability development: A ChatGPT perspective
models against human performance would provide nov. J., 1(2), 97–104. doi: [Link]
important insights into the relative sustainability of zenodo.7855594.
AI development. Collaborative working with busi- Jiafu, A. N., Wenzhi, D. I. N. G., and Chen, L. I. N. (2023).
ness stakeholders, AI developers, and environmen- Correspondence: ChatGPT: Tackle the growing car-
bon footprint of generative AI. Nature, 615(7953),
tal specialists would promote a multidisciplinary
586. doi:[Link]
approach to comprehending and reducing the envi-
2.
ronmental effects of AI models. These collaborations Tanushree, K. (2023). Is ChatGPT harmful to the environ-
might make it easier for people to acquire informa- ment. Sigma Earth. [Link]
tion, share expertise, and create sustainable practices harmful-to-the-environment/.
and norms. Mehreen, K. and Chaudhry, M. N. (2023). Artificial intel-
As time goes on carrying out longitudinal studies ligence and the future of impact assessment. Available
over a lengthy period of time, researchers will be able at SSRN 4519498. doi: [Link]
to monitor changes in the environmental effect of AI ssrn.4519498.
models. In order to lessen the environmental impact Ali, K. S., Khuwaja, P., and Dev, K. (2023). ChatGPT needs
of AI models, this would assist to detect trends, tech- SPADE (Sustainability, PrivAcy, Digital divide, and
Ethics) evaluation: A review. arXiv preprint arX-
nological improvements, and viable mitigation tech-
iv:2305.03123. doi: [Link]
niques. Adding case studies and field research to
iv.2305.03123.
secondary data analysis will offer insightful informa- Alexandra Sasha, L., Viguier, S., and Ligozat, A.-L. (2022).
tion on the precise environmental effects of AI mod- Estimating the carbon footprint of bloom, a 176b
els in various industries or applications. These studies parameter language model. arXiv preprint arX-
could concentrate on actual situations and take into iv:2211.02001.
account things like the setup of data centers, power Lund, B. D., Wang, T., Mannuru, N. R., Nie, B., Shimray, S.,
sources, and energy-saving techniques. and Wang, Z. (2023). ChatGPT and a new academic
reality: Artificial Intelligence-written research papers
and the ethics of the large language models in schol-
References arly publishing. J. Assoc. Inform. Sci. Technol., 74(5),
Gulzar, A., Ihsanullah, I., Naushad, M., and Sillanpää, M. 570–581. doi:[Link]
(2022). Applications of artificial intelligence in water Khattar, N., Singh, J., and Sidhu, J. (2020). An energy ef-
treatment for optimization and automation of adsorp- ficient and adaptive threshold VM consolidation
tion processes: Recent advances and prospects. Chem. framework for cloud environment. Wireless Personal
Engg. J. 427, 130011. doi:[Link] Comm., 113, 349–367.
cej.2021.130011. Sophie, M. (2023). The environmental impact of ChatGPT:
An, J., Ding, W., and Lin, C. (2023). ChatGPT: Tackle A call for sustainable practices in AI development.
the growing carbon footprint of generative AI. Na- [Link]. [Link]
ture, 615(7953), 586. doi:[Link] chatgpt/#:~:text=The/Environmental/Impact/of/Data/
d41586-023-00843-2. Centres&text=According/to/estimates/C/ChatGPT/
Priyanka, B. and Sharma, K. D. (2022). A critical insight emits.
into the role of artificial intelligence (AI) in tour- Kannan, N. (2023). AI-enabled water management sys-
ism and hospitality industries. Pacific Business Rev. tems: An analysis of system components and interde-
Int., 15(3), 76–85. [Link] pendencies for Wwater conservation. Eigenpub Rev.
publication/371110624_A_Critical_Insight_into_the_ Sci. Technol., 7(1), 105–124. [Link]
Role_of_Artificial_Intelligence_AI_in_Tourism_and_ com/[Link]/erst.
Hospitality_Industries. David, P., Gonzalez, J., Hölzle, U., Le, Q., Liang, C., Mun-
Biswas, S. S. (2023). Potential use of chatGPT in global guia, L.-M., Rothchild, D., So, D. R., Texier, M., and
warming. Ann. Biomed. Engg., 51(6), 1126–1127. Dean, J. (2022). The carbon footprint of machine
doi:[Link] learning training will plateau, then shrink. Computer,
David, D. (2023). AI creates new environmental injus- 55(7), 18–28. doi: 10.1109/MC.2022.3148714.
tices, but there’s a fix. News. [Link] Chengwei, Q., Zhang, A, Zhang, Z., Chen, J., Yasunaga,
articles/2023/07/12/ai-creates-new- environmental- M., and Yang, D. (2023). Is ChatGPT a general-
injustices-theres-fix. purpose natural language processing task solv-
Euronews. (2023). Chat: What is the carbon footprint er?. arXiv preprint arXiv:2302.06476. doi:[Link]
of generative AI models? Euronews. [Link] org/10.48550/arXiv.2302.0647.
[Link]/next/2023/05/24/chatgpt-what- Partha Pratim, R. (2023). AI-Assisted Sustainable Farming:
is-the-carbon-footprint-of-generative-ai-mod- Harnessing the Power of ChatGPT in Modern Agri-
els#:~:text=Researchers%20estimated%20that%20 cultural Sciences and Technology. ACS Agricultural
creating%20GPT. Science & Technology. 460–462. Doi: [Link]
George, A. S., George, A. S. H., and Martin, A. S. G. (2023). org/10.1021/acsagscitech.3c00145
The environmental impact of AI: A case study of wa- Preeti, S. and Priyanka, P. (2019). Climate change and sus-
ter consumption by Chat GPT. Partn. Univ. Int. In- tainable development: Special context to Paris agree-
62 Environment and sustainability development: A ChatGPT perspective
ment. Proc. Int. Conf. Sustain. Comput. Sci. Technol. Aimee, V. W. (2021). Sustainable AI: AI for sustainabil-
Manag. (SUSCOM). doi:[Link] ity and the sustainability of AI. Spring Link, 1(3),
ssrn.3356829. 213–218. doi:[Link]
Sohail, S. S., Farhat, F., Himeur, Y., Nadeem, M., Madsen, 00043-6.
D. Ø., Singh, Y., Atalla, S., and Mansoor, W. (2023). Demetris, V., Chaudhuri, R., and Chatterjee, S. (2023).
Decoding ChatGPT: A taxonomy of existing research, Role of ChatGPT and skilled workers for business
current challenges, and possible future directions. J. sustainability: Leadership motivation as the mod-
King Saud Univ. Comp. Inform. Sci.. 101675. doi: erator. Sustainability, 15(16), 12196. doi: [Link]
[Link] org/10.3390/su151612196.
Aman, S., Jain, S., Maity, R., and Desai, V. R. (2022). De- Pooja, Y. (2023). Explained: What is the water footprint
mystifying artificial intelligence amidst sustainable ag- of AI and how AI tools are raising environmental
ricultural water management. Curr. Dir. Water Scar. concerns. IndiaTimes. [Link]
Res., 7, 17–35. doi:[Link] explainers/news/explained-what-is-the-water-foot-
323-91910-4.00002-9. print-of-ai-and-how-ai-tools-are-raising-environmen-
Timm, T., Flath, C. M., Weinhardt, C., van der Aalst, W., [Link].
and Hinz, O. (2023). Welcome to the era of chat- Jun-Jie, Z., Jiang, J., Yang, M., and Ren, Z. J. (2023).
gpt. The prospects of large language models. Busin. ChatGPT and environmental research. Envi-
Inform. Sys. Engg., 65(2), 95–101. doi: [Link] ron. Sci. Technol. doi:[Link]
org/10.1007/s12599-023-00795-x. showCitFormats?doi=10.1021/[Link].3c01818.
9 GAI in healthcare system: Transforming research in
medicine and care for patients
Mahesh A.1,a, Angelin Rosy M.2, Vinodh Kumar M.3, Deepika P.4,
Sakthidevi I.5 and Sathish C.6
1,4
Sri Sairam College of Engineering, Karnataka, India
2,6
Er. Perumal Manimekalai College of Engineering, Tamilnadu, India
3
P.S.V College of Engineering and Technology, Tamilnadu, India
5
Adhiyamaan College of Engineering, Tamilnadu, India
Abstract
GAI also known as generative artificial intelligence, represents a category of artificial intelligence (AI) that possesses the
capability to produce novel content, encompassing images, written text, and music. Although it remains in its emerging
phases of advancement, this technology holds the promise of revolutionizing numerous sectors, ranging from healthcare and
finance to entertainment. The subject of GAI is rapidly developing and holds the capacity to transform the field of health-
care. The adoption of GAI technology has revolutionized the healthcare industry, transforming the way patients are treated
and medical research is conducted. This article explores the many potential applications of GAI in healthcare, including its
ability to improve diagnostic accuracy, optimize treatment, accelerate drug discovery, and enhance medical image analysis.
GAI, as demonstrated by advanced neural network algorithms like variational autoencoders (VAEs) and generative adver-
sarial networks (GANs) enables healthcare practitioners, medical analyst, technologist and scientists to generate realistic
and high-fidelity medical data. Using this technology, medical professionals can improve diagnosis accuracy by combining
varied information about patients, allowing for more robust and individualized treatment strategies. Furthermore, GAI aids
in the generation of realistic medical images, allowing medical practitioners to better grasp and interpret difficult illnesses.
In the field of drug exploration, GAI speeds up the process for determining possible compounds and molecules, saving time
and money over traditional methods. It investigates how GAI encourages interaction among human experts and artificially
intelligent machines, allowing medical practitioners to make better decisions. This complementary partnership takes use of
the capacity of artificial intelligence to analyze large datasets, detect trends, and recommend viable treatment paths, while
human knowledge provides the context-sensitive knowledge required for informed decision-making. Ethical concerns and
obstacles related with the application of GAI to medical procedures are also addressed, with an emphasis on the importance
of responsible application, data protection, and transparency. The healthcare sector aspires ready to bring in a new era of
distinctive effective and cost-effective treatment for patients and research in medicine by adopting the revolutionary potential
of GAI and managing its ethical consequences.
Keywords: Generative artificial intelligence, healthcare, generative adversarial networks, variational autoencoders, ethical
a
sathishinfy@[Link]
64 GAI in healthcare system: Transforming research in medicine and care for patients
generator improves over time at producing data that in the upcoming years thanks to the ongoing develop-
is accurate. ment of this technology (Singh et al., 2019; Jovanović
et al., 2022; Samant et al., 2022).
b. Variational autoencoders (VAEs)
VAEs are a type of neural network that can compress II. Evolution of GAI
data into a smaller, hidden space, and then decom-
press it back to the original data. The latent space is a The advent of GAI signifies resulted in substantial
space with fewer dimensions that captures the data’s advances in ML and AI. In the following sections,
basic characteristics. After learning to represent data a timeline of how generative AI has progressed is
in the space known as the latent space, the VAE can discussed.
be used to produce new data by sampling points from
the latent space and decoding these again into the a. Beginning principles (1950–1990s)
data that was originally collected space. During the early years of AI research, the underlying
principles of GAI were established. In domains such
c. Recurrent neural networks (RNNs) as natural language processing and music creation,
RNNs are a type of neural network system that is researchers investigated rule-based systems and sym-
capable of processing sequential data. As a result, bolic representations to generate content.
they are well-suited for generating text, music, and
other sorts of data with a periodic order. b. The rebirth of neural networks in the 2000s
The various approaches needed to build a GAI sys- The “deep learning revolution,” or the resurrection of
tem will vary depending on the application. GANs, artificial neural networks, was critical in the creation
for instance, are frequently used to produce images, of generative artificial intelligence. Neural networks
but VAEs are frequently used to generate text. The with deep learning revealed the ability to learn data
following are some of the steps involved in developing hierarchies, allowing for the development of higher-
a GAI system: level and more intricate outputs (Davies et al., 2021).
i. Data collection: The initial step involves gather- c. Variational autoencoders (VAEs) (2013)
ing an extensive array of data pertinent to the VAEs pioneered a probabilistic approach to genera-
intended application. For instance, if the goal is tive modeling. To construct a latent space represen-
to create cat photographs, the initial task entails tation of data, they incorporated aspects from both
amassing a dataset comprising images of cats. generative and recognition models. VAEs enabled
ii. Dataset pre-processing: Before the data can be seamless interpolation and manipulation of latent
employed for model training, it might necessitate space data points.
pre-processing. This step could encompass tasks
such as data cleansing, noise reduction, and nor- d. Generative adversarial networks (GANs) (2014)
malization. GANs, suggested by Ian Goodfellow and colleagues,
iii. Algorithm choice: Multiple models exist for con- represented a significant development in generative
structing a GAI system. The selection of an ap- AI. GANs are made up of the discriminator and gen-
propriate model hinges on the specific applica- erator, two distinct neural networks that compete
tion and the quantity of available data. with one another in a manner akin to a game. While
iv. Model training: The algorithm is subjected to a the discriminator seeks to distinguish between genu-
training process using a dataset. The duration of ine and produced data, the generation process aims to
this process can vary based on the dataset’s size provide data that is as realistic as possible.
and the desired precision of the model, some-
times spanning a considerable timeframe. e. Visualizing the future era (2014–current)
v. Generate fresh data: Following the completion During this period, GAN gained prominence due
of model training, it becomes feasible to employ to their ability to create high-resolution images that
the model for generating novel data. The ap- closely resemble authentic photographs. Renowned
proach employed in this process is contingent on GAN architectures like deep convolutional GAN
the specific framework being used and dictates (DCGAN), StyleGAN, and BigGAN elevated the
the manner in which the new data is crafted. caliber and diversity of the generated images to new
heights (Guo et al., 2022).
While GAI systems are currently in their nascent
stages of research, they hold immense potential to f. Text and language making (2015–current)
revolutionize various industries we may anticipate Progress in the realm of natural language process-
seeing more cutting-edge and significant uses of GAI ing and deep learning has yielded the creation of text
Applied Data Science and Smart Systems 65
and language generative frameworks. Innovations produce offensive content, or reproduce biases exist-
like long- short- term memory (LSTM) networks and ing in training data sparked debate over ethical imple-
transformers have empowered the generation of logi- mentation and mitigating techniques.
cally connected and contextually fitting textual con-
tent (Sathish et al., 2023). j. Ongoing exploration and advancement (current and
beyond)
g. Music and audio production (2016–current) Researchers are working on ways to improve the
Generative algorithms have been used to compose quality of generated content by making it more
music and synthesize audio. Melodies, harmonies, realistic, diverse, and controllable. Hybrid models
and even full music recordings have been generated that combine different GAI techniques are being
using recurrent neural networks along with differ- developed to improve the performance of models.
ent sequence-to-sequence algorithms. WaveGAN and Creative applications of GAI are being explored
other approaches have also showed promise in pro- in areas such as art, music, and video games.
ducing realistic signals for audio. Interdisciplinary collaborations between researchers
from different fields are helping to advance the state
h. Applications in healthcare and science (2010–cur- of GAI research.
rent) Advances in deep learning architectures, computa-
GAI has been used in healthcare since 2010s for a tional power, and the availability of massive datasets
variety of purposes, including image interpretation, have driven the development of GAI. As technologi-
drug development, and personalized medicine. GANs cal advancements continue, AI with generative capa-
and VAEs have found use in creating artificial medical bilities has the potential to impact a wide range of
images to improve diagnostic accuracy and expand persistence, from entertainment and art to health-
small datasets (Figure 9.1) (Rebecca Perkins et al., care and scientific research in this system (Cai et al.,
2022). 2019).
studied is the first step in building a generative AI sys- Autoregressive algorithms generate information
tem. In the following sections, the stages below out- throughout a sequential manner, projecting the next
line the general technique for developing a generative component based on prior components.
artificial intelligence system.
d. Functions of loss
a. Select a generative approach GANs – The generator and discriminator networks
Variational autoencoders, GANs, and autoregressive are adversarial trained. The discriminator strives to
models such as transformers are examples of genera- accurately classify both real and generated data, while
tive models that can be used. As per an individual the generator aims to diminish the discriminator’s
wish, he/she can choose the model that best fits the ability to distinguish genuine from generated data.
data which has to be developed (Figure 9.2). VAEs – To guarantee space of latent information is
well-structured, the model is trained using a combina-
b. Collection of data and pre-processing tion of reconstruction loss (how well the generated data
A wide and representative dataset of the type of data matches the original input) and a regularization term.
(as per wish) is compiled to generate (e.g., photo- Autoregressive models optimize the expected prob-
graphs, document, audio, etc.) (Hajarolasvadi et al., ability distribution over the next element using nega-
2019). tive log-likelihood loss.
The data is pre-processed to ensure that it remains
consistent and in the correct format for the mathemat- e. System training
ical framework of choice. This could include scaling The representation’s parameters are prepared infor-
photos, standardizing the values of pixels, represent- mally. The model is feed with real data (for GANs,
ing text, and so on. this is the discriminator’s input; for VAEs, this is the
input for encoding and reconstruction). Fake data
c. Design of architecture samples are generated and feed to the model.
The design of GANs comprises of a generating net- The loss for both the real and generated data is
work and a discriminator network. The genera- calculated and use back propagation to update the
tor generates samples of data, and a discriminator model’s parameters (Walczak et al., 2018).
attempts to differentiate between genuine and pro-
duced samples. f. The iteration process and refinement
The design for VAEs consists of an encoder net- Continuous training of a computational framework
work, a decoder network, and a latent space in involves multiple rounds of iterative adjustments
between. The encoder converts input data to a lower- to enhance its performance. To avoid over fitting,
dimensional latent space, from which the decoder keep an eye on the model’s results on the validation
produces data. information.
g. Assessing and further refinement fitting. Nearly 10% of the data is used to gauge the
To analyze the quality and diversity of the gener- effectiveness of the model.
ated samples, domain-specific metrics or judgment
by humans are used. To improve outcomes, tweak c. Pre-processing of data
hyper parameters, model architectural design, as well Pictures are resized to a common resolution. Scale
as training procedures as appropriate (Alam et al., pixel values to a standard range, such as 0 to 1. To
2018). boost dataset diversity, supplement data with rota-
tions, flips, and other transformations.
h. Development / The next generation
Following the training session, the generator’s results d. Selection of appropriate model
can be employed to produce new data samples by A convolutional neural network (CNN) architecture
supplying random deep space points (for VAEs) or is choose that is appropriate for image categorization
randomly generated noise vectors (for GANs). when selecting a model.
Creating generative AI models for healthcare that Chatbots for personalized medical care
are both efficient and secure by adhering to these thor- Medical chatbots can be developed by healthcare
ough material and methodology standards, which will institutions to give patients with tailored medical
also help to improve patient care and outcomes. It is information and suggestions. Babylon healthcare, for
kept in mind that successful development of health- illustration, has created a chatbot that uses GAI to ask
care AI necessitates a multi-disciplinary approach and patients about the symptoms they are experiencing
collaboration with healthcare professionals. and provide individualized medical recommendations.
capturing data, producing electronic health records, personalized patient care. However, the challenges
and reducing complex medical terminology for and ethical considerations inherent in integrating
patient comprehension (Figure 9.3). generative AI within healthcare must be thoroughly
deliberated upon. With ongoing exploration in addi-
VI. Challenges generative AI in healthcare tion improvement, potential for generative artifi-
cial intelligence to reshape healthcare and amplify
Although AI that regenerates has enormous potential patient well-being in the coming years remains
in healthcare, several problems must be overcome. substantial.
1
Govt College for Women, Shahzadpur (Ambala), Haryana, India
4
Govt College, Naraingarh (Ambala), Haryana, India
Abstract
This paper is the fuzzy analysis of a queue network model with the assumptions of pre-emptive priority discipline on biserial
subsystems and general arrival is on parallel subsystems. It is presupposed that service time and the interval between two
succeeding arrivals follow the Poisson distribution. Both arrivals and service costs are fuzzy in nature. Performance of the
purposed model evaluated by using L-R fuzzy numbers. L-R method is more flexible and simplest method for fuzzy analysis
as compared to other existing methods. Fuzzy triangular number and all classical formulae used to calculate fuzzy queue
characteristics. A numerical calculation well illustrated the results.
Keywords: Fuzzy number, priority, parallel channels, biserial server, L-R method
a
aartisaini195@[Link]
72 Fuzzy L-R analysis of queue network with priority
To solve the above equation, by applying L’ hos- = 1. And find the utilization factors at different servers
pital rule with conditions | Z1|=| Z2|=| Z3|=| in stochastic environment
Z4|=|Z5|=|Z6|=|Z7|=1 and H(Z1, Z2, Z3, Z4, Z5, Z6, Z7)
Applied Data Science and Smart Systems 73
Fuzzified model
With the conditions exist if γ1, γ2, γ3, γ4, γ5, γ6, γ7 ≤ 1
Let us represent approximate crisp parameters
IV. Numerical illustration in the form of fuzzy numbers as
For particular crisp values, we get then
Using Table 10.1, utilization factor and queue from stochastic environment results, the fuzzy utiliza-
length is tion factor and queue characteristics can be written as
74 Fuzzy L-R analysis of queue network with priority
VII. Results
Fuzzy lengths of queues • Utilization of first server by high priority cus-
tomers lies between 0.1504 and 0.8271. Utiliza-
tion factor and partial queue length’s maximum
allowed values are 0.3906 and 0.6410. Utiliza-
tion of first server by low priority customers lies
between 0.2061 and 1.5529. The partial queue
lengths and Utilization factor maximum allowed
Average waiting time
values are 1.6738 and 0.6260.
• Utilization of second server by high priority
customers lies between 0.1772 and 0.8717.
Utilization factor and partial queue length’s
maximum allowed values are 0.4167 and
VI. Numerical illustration 0.7144. Utilization of second server by low
Table 10.3 is the fuzzy L-R representations of fuzzy priority customers lies between 0.2695 and
triangular numbers from Table 10.2. 1.6776. Utilization factor and partial queue
Using these numerical values, we get L-R represen- length’s maximum allowed values are 0.7045
tations of traffic intensity at servers are, and 2.3841.
• Utilization of third server lies between 0.2679
and 1.2498. Utilization factor and partial queue
length’s maximum allowed values are 0.5555 and
1.2497.
• Utilization of fourth server lies between .084 and Mittal, Meenu, T. P. Singh, and Deepak Gupta. (2015).
1.2821. Utilization factor and the length of par- Threshold effect on a fuzzy queue model with batch
tial queue maximum potential values are 0.4 and arrival. Arya Bhatta Journal of Mathematics and In-
0.6. formatics. 7(1), 109–118.
Singh, T. P., Mittal, M., and Gupta, D. (2016). Modelling of
• Utilization of fifth server lies between 0.2153
a bulk queue system in triangular fuzzy numbers using
and 1.4258. Utilization factor and partial queue
α-cut. Int. J. IT Engg., 4(9), 72–79.
length’s maximum potential values are 0.6087 Devaraj and Jayalakshmi. (2012). A fuzzy approach to pri-
and 3.4984. ority queues. Int. J. Fuzzy Math. Sys., 2(4), 479–488.
Gupta, D., Sharma, S., and Gulati. (2011). On steady state
VIII. Conclusion behavior of a network queuing model with bi-serial
and parallel channels linked with a common server.
In the present work, based on L-R fuzzy arithmetic Comp. Engg. Intel. Sys., 2(2).
operations, priority queues have been analyzed by Seema, Gupta, D., and Sharma, S. (2013). Analysis of bise-
the L-R technique. This method is used to evaluate rial servers linked to a common server in fuzzy envi-
numerical values of various performances of queues ronment. Int. J. Comp. Sci. Math., 68(6), 26–32.
like traffic intensity and length of queues at different Sharma, S., Gupta, D., and Seema. (2015). Network Aanaly-
servers in fuzzy environment. Fuzzy L-R representa- sis of fuzzy bi-serial and parallel servers with a multi-
stage flow shop model. 21st Int. Cong. Model. Simul.
tion is more informative than basic classical methods
Gold Coast Australia, 697–703.
in stochastic environment. For this numerical calcu-
Kalpana. (2021). Evaluation of performance measures of
lation is used to authenticate the study. While using fuzzy queues with preemptive priority using different
same approximate crisp and fuzzy data, then deter- fuzzy numbers. Adv. Appl. Math. Sci., 20(11), 2975–
mine the outcomes in the event of precise numbers for 2985.
the fraction of both type customers high and low pri- Kalpana and Anusheela. (2018). Analysis of fuzzy non-pre-
ority using the first server are 39.06% and 75.68%, emptive priority queue using non-linear programming.
second server usage by high and low priority custom- Int. J. Math. Trends Technol. (IJMTT), 56(1), 71–80.
ers is 50% and 85.98%, third, fourth and fifth server Selvakumaria and Revathi (2021). Analysis of fuzzy non-
usage are 66.66%, 28.85% and 77.77%, respectively. preemptive priority queuing model with unequal ser-
Accessing these servers while dealing with ambiguous vice rate. Turkish J. Comp. Math. Edu., 12(5), 1457–
1460.
data is 39.06% and 62.60%, 41.67% and 70.45%,
Rita, W. and Robert. (2009). Application of fuzzy set theory
55.55%, 40% and 60.87%, respectively. Thus, from
to retrial queues. Int. J. Algorith. Comput. Math., 2(4),
results we can observe that utilization of 1st and 2nd 9–18.
server by high priority customers is approximate same Ritha, W. and Menon, S. B. (2011). Fuzzy n policy queues
but utilization of servers in stochastic environment by with infinite capacity. J. Phy. Sci., 15, 73–82.
low priority customers is high as compared to fuzzy Ning and Zhao. (2009). Analysis on random fuzzy queu-
environment. The usage of 4th sever in crisp data is ing systems with finite capacity. 9th Int. Conf. Elec.
28.85% and in fuzzy data is 40%. Thus, the study in Busin., 1–7.
future can be extended for more queuing models with Yang, W. and Li. (2010). Fuzzy analysis for the n-policy
batch arrival, priority arrivals on parallel subsystem queues with infinite capacity. Int. J. Inform. Manag.
and biserial servers instead of parallel subsystem. Sci., 21, 41–45.
Srinivasan. (2013). Fuzzy queuing model using DSW algo-
rithm. Int. J. Adv. Res. Math. Comp. Appl., 1(1), 1–6.
References Ritha, W. and Vinnarasi, J. S. (2017). Analysis of priority
queuing models: L-R method. Ann. Pure Appl. Math.,
Prade, (1980). An outline of fuzzy or possibilistic mod-
15(2), 271–276.
els for queuing systems. Wang P. P. and Chang S. K.
Mukeba, J. P., Mabela and Ulungu. (2015). Computing
(eds), Fuzzy Sets. Plenum Press. 147–153. [Link]
fuzzy queuing performance measures by L-R method.
org/10.1007/978-1-4684-3848-2_13.
J. Fuzzy Sets Valued Anal., 1, 57–67.
Li and Lee. (1988). Analysis of fuzzy queues. Proc. NAFIPS,
Mukeba, J. P. (2016). Application of L-R method to single
158–162.
server fuzzy retrial queue with patient customers. J.
Li and Lee. (1989). Analysis of fuzzy queues. Comp.
Pure Appl. Math. Adv. Appl., 16(1), 43–59.
Math. Appl., 17(7), 1143–1147. [Link]
Saini, V., Gupta, D., and Tripathi, A. K. (2022). Analysis
org/10.1016/0898-1221(89)90044-8.
of heterogeneous feedback queue model in stochastic
Li, K. and Chen. (1999). Parametric programming to the
and in fuzzy environment using L-R method. Math.
analysis of fuzzy queues. Fuzzy Sets Sys., 107, 93–100.
Stat., 10(5), 918–924.
http:/[Link]/10.1016/S0165- 0114(97)00295-9.
Singh, T. P., Kusum, and Gupta, D. (2010). On network
queue model centrally linked with common feedback
channel. J. Math. Sys. Sci. 6(2), 18–31.
11 Blood bank mobile application of IoT-based android
studio for COVID-19
Basetty Mallikarjuna1, Sandeep Bhatia2,a, Neha Goel3, and
Bharat Bhushan Naib4
1
Department of Information Technology, Institute of Aeronautical Engineering, Dundigal-500043, Tamil Nadu, India
2,4
School of Computing Science and Engineering, Galgotias University Greater Noida, Uttar Pradesh, India
3
Department of Electronics & Communication Engineering, RKGIT, Ghaziabad, India
Abstract
It is impossible to manufacture the blood, as it can be given by the donors. Blood bank retrieval information can be given
through the android studio application, but there is not much work on the integrated environment like IoT sensor connected
with android studio application development. This paper provides the IoT healthcare sensors connected to the android
studio mobile application development blood donors and blood receivers. The mobile application is most useful in an
emergency during the COVID-19 pandemic. This observational study gives the web-based application development and also
android-studio mobile application development for blood bank information retrieval system. The results are carried out in a
real-time environment and updated features of the blood bank mobile application.
Keywords: Internet of things, COVID-19, blood bank, mobile application, android studio
sandeepbhatia1711@[Link]
a
Applied Data Science and Smart Systems 77
Table 11.1 Frequency of occurring in different blood blood information management mobile application
groups (Fahim et al., 2016). which has its mobile search engine used to search for
blood donors and receivers from the registered appli-
S. No Approximate frequency of occurring blood type
cation. This study also provides that registered users
1 O +ve: 1 person might be among 3 persons send a notification to donors and receivers. The pro-
2 A +ve: 1 person might be among 3 persons posed application also has certain disadvantages, it
requires an internet connection and manages particu-
3 O -ve: 1 person might be among 15 persons
lar functions required for a large database.
4 A -ve: 1 person might be among 16 persons In this paper section 2 deals with the related work
5 B +ve: 1 person might be among 12 persons existing to differentiate the proposed methodology,
6 B -ve: 1 person might be among 67 persons section 3 deals with the methodology, section 4 pro-
7 AB +ve: 1 person might be among 29 persons vides the implementation, and section 5 deals with the
8 AB -ve: 1 person might be among 167 persons
conclusion followed by references.
The user can also search for the blood in any city by
clicking on the search button at the top and provid-
ing the details like which blood group the user wants
and in which city and it will show you the details if
anybody was there. The user can also become a donor
so the user can save any life by giving the blood group Figure 11.3 Snapshot of the API of android studio
and people can see who is given the blood in that city.
The activity is divided into several parts like login
activity to increase the security of the application.
Then the main activity tells about the people need-
ing the blood and also the activity where the user can
request the blood or become a donor. So, it’s a very
modern compact and the best application to over-
come a serious problem as shown in Figure 11.4.
One of the significant and important of this obser-
vational study is not alert the blood and also provides
the various healthcare issues of the registered users
during the COVID-19 pandemic.
V. Implementation
To create a simple android application project, to set
up the application for the following steps below:
File -> New -> Select New.
Fill in all the entries shown in the above Figure
11.3. Set the name and location of the project. Select
the language in which you want to code. After fill-
ing all the fields click on finish. Once the project is
Figure 11.4 GUI of new app
successfully created the screen will show as shown in
Figure 11.4.
There are some directories and files in the android pandemic. Here are some examples of how the appli-
project which we should be created before start- cation has been used in the fight against COVID-19:
ing our application as shown in Figure 11.5 and the
description of packages as shown in Table 11.2. • Blood donation and distribution in real-time
• Contactless donation
• Inventory management
VI. Applications of our work
• Emergency response
The android studio-based blood bank mobile appli- • Analysis of data to predict demand.
cation was created utilizing IoT-based technology • Remote health monitoring.
to handle COVID-19 difficulties, and it can have a • Donation of post-recovery blood plasma
variety of uses and advantages in the context of the • Community awareness and involvement
80 Blood bank mobile application of IoT-based android studio for COVID-19
conditions, lowering waste, and increasing overall in Computational Intelligence and Communication
effectiveness. Technology: Proceedings of CICT, 2019: 501–510.
Incorporating these future scopes would not only doi: [Link]
increase the blood bank mobile application’s efficiency Mufaqih, Sukron, Abiyyu Fawwaz Kanz, Sahid Nur Rama-
dhan, and Ahmad Nurul Fajar. (2020). Blood Bank
and efficacy but will also make a major improvement
Information System Based on Cloud In Indonesia.
to healthcare services generally, which is especially
IOP Conf. Series: Journal of Physics: Conf. Series
important during pandemics like COVID-19. 1179(2019) 012028. 1–6.
Singh, Jaiteg, Gaurav Goyal, and Rupali Gill. (2020). Use of
References neurometrics to choose optimal advertisement meth-
od for omnichannel business. Enterprise Information
Al Bassam, N., Hussain, S. A., Al Qaraghuli, A., Khan,
Systems. 14(2): 243–265. doi: [Link]
J., Sumesh, E. P., and Lavanya, V. (2021). IoT based
/17517575.2019.1640392
wearable device to monitor the signs of quarantined
Sultanul, A. and Taposi, S. (2021). Blood bank mobile ap-
remote patients of COVID-19. Informat. Med. Un-
plication. 1–39.
lock., 24, 100588.
Mahima, N., Nigam, C., and Chaurasia, N. (2019). m-Health:
Aderonke Anthonia, K., Adeniyi, A. E., Ogundokun, R. O.,
community-based android application for medical ser-
and Ochigbo, S. A. (2019). An android based blood bank
vices. Smart Healthcare Sys., 69–81. CRC Press.
information retrieval system. J. Blood Med., 119–125.
Pohandulkar, Surabhi, S., and Khandelwal, C. S. Blood
Aman, S., Shah, D., Shah, D., Chordiya, D., Doshi, N., and
bank app using raspberry PI (2018). 2018 Int. Conf.
Dwivedi, R. (2022). Blood bank management and in-
Comput. Tech. Elec. Mech. Sys. (CTEMS), 355–358.
ventory control database management system. Proce-
Reddy, C. K. K., Anisha, P. R., and Prasad, L. N. (2016). A
dia Comp. Sci., 198, 404–409.
novel approach for detecting the bone cancer and its
Priya, P., V. Saranya, S. Shabana, and Kavitha Subramani.
stage based on mean intensity and tumor size. Recent
(2014). The optimization of blood donor information
Res. Appl. Comp. Sci., 20(1), 162–171.
and management system by Technopedia. Interna-
Narasimha, P. and Munirathnam Naidu, M. (2013). Gain
tional Journal of Innovative Research in Science, En-
ratio as attribute selection measure in elegant decision
gineering and Technology. 3(1), 1–6.
tree to predict precipitation. 2013 8th EUROSIM
Muhammad, F., Cebe, H. I., Rasheed, J., and Kiani, F.
Cong. Model. Simul., 141–150.
(2016). mHealth: Blood donation application using
Tatale, Subhash, and V. Chandra Prakash. (2020). Enhanc-
android smartphone. 2016 Sixth Int. Conf. Dig. In-
ing acceptance test driven development model with
form. Comm. Technol. Appl. (DICTAP), 35–38.
combinatorial logic. International Journal of Ad-
Ayman, A., Mallikarjuna, B., Saudagar, A. K. J., Sharma, M.,
vanced Computer Science and Applications., 11(10),
and Poonia, R. C. (2022). Improvement of automatic
268–278.
glioma brain tumor detection using deep convolution-
Sastry, J. K. R. and Lakshmi Prasad, M. (2019). Testing
al neural networks. J. Comput. Biol., 29(6), 530–544.
embedded system through optimal mining technique
Krishna, P. V., Gurumoorthy, S., Obaidat, M. S., Mallikarjuna,
(OMT) based on multi-input domain. Int. J. Elec.
B., and Arun Kumar Reddy, D. (2019). Healthcare appli-
Comp. Engg., 9(3), 2141–2150.
cation development in mobile and cloud environments.
Prasad, M. L. and Sastry, J. K. R. (2018). Generation of
Internet of things Personal. Healthcare Sys., 93–103.
test cases using combinatorial methods based multi-
Archit, S., Mallikarjuna, B., Murtuza, M., and Tiwari, V.
output domain of an embedded system through the
(2021). Design and implementation of superstick for
process of optimal selection. Int. J. Pure Appl. Math.,
blind people using internet of things. 2021 3rd Int. Conf.
118(20), 181–189.
Adv. Comput. Comm. Con. Netw. (ICAC3N), 691–695.
Sandeep, B., Mallikarjuna, B., Gautam, D., Gupta, U., Ku-
Mallikarjuna, B., Sathish, K., Gitanjali, J., and Venkata
mar, S., and Verma, S. (2023). The Future IoT: The
Krishna, P. (2021). An efficient vote casting system
current generation 5G and next generation 6G and
with aadhar verification through blockchain. Int. J.
7G technologies. 2023 Int. Conf. Dev. Intel. Comput.
Sys. Sys. Engg., 11(3–4), 237–256.
Comm. Technol. (DICCT), 212–217.
Mallikarjuna, B. (2022). Feedback-based resource utili-
Ganai, P. T., Bag, A., Sable, A., Abdullah, K. H., Bhatia, S.,
zation for smart home automation in fog assistance
and Pant, B. (2022). A detailed investigation of imple-
IoT-based cloud. Res. Anthol. Cross-Dis. Des. Appl.
mentation of internet of things (IOT) in cyber security
Automat., 803–824.
in healthcare sector. 2022 2nd Int. Conf. Adv. Com-
Khan, Mohammad Asaduzzaman, Hasibur Rahaman, Iske-
put. Innov. Technol. Engg. (ICACITE), 1571–1575.
daheer Alam, Khayrul Alam, Sumon Mondal, and
Sandeep, B., Goel, N., and Verma, S. (2023). The current
Alimuzzaman Khan. (2021). Development of Applica-
generation 5G and evolution of 6G to 7G technolo-
tion to Find A Nearby Live Blood Donor Using the
gies: The future IoT. Handbook Res. Mac. Learn-En-
Updated Location e-Information. International Jour-
abled IoT Smart Appl. Across Indust., 456–478.
nal of Electrical Engineering and Applied Sciences
Sandeep, B., Goel, N., Ahlawat, V., Naib, B. B., and Singh, K.
(IJEEAS), 4(1), 17–21.
(2023). A comprehensive review of IoT reliability and
Modi, Nandini, and Jaiteg Singh. (2021). A review of various
its measures: Perspective analysis. Handbook Res. Mac.
state of art eye gaze estimation techniques. Advances
Learn-Enabled IoT Smart Appl. Across Indust., 365–384.
12 Selection of effective parameters for optimizing software
testing effort estimation
Vikas Chahara and Pradeep Kumar Bhatia
Guru Jambheshwar University of Science & Technology Hisar, India
Abstract
Software testing holds a significant role within the realm of software development. Its purpose is to bolster and elevate the
reliability and quality of software. This encompassing process involves several key steps, including estimating the required
testing effort, assembling an appropriate test team, formulating effective test cases, carrying out software execution using
these test cases, and meticulously analyzing the outcomes derived from these executions. Thus, precise software testing effort
estimation holds high significance and governs the overall cost of the software development. To support accurate estimation
of software testing effort, the paper presents a detailed analysis and categorization of various factors have great impact on
the software testing effort. The analysis shows that the parameters include various elements such as quality, stability, risk,
resources, etc., that have a significant impact on software testing effort. Thus, the paper contributes towards a platform for
extracting the basic information essential for supporting software testing effort estimation for successful project planning
and execution.
[Link]@[Link]
a
Applied Data Science and Smart Systems 83
ternal files, and external interfaces to determine defects, vulnerabilities, and inconsistencies in the
the complexity and size of a software system. software before it’s released to end-users (Sharma and
This metric is often used in software estimation, Kushwaha, 2011; Hidmi and Sakar, 2017; Brar et al.,
project management, and cost analysis. Math- 2022). The effort invested in software testing is influ-
ematically it can be expressed as: enced by several parameters that impact the complex-
ity and scope of the testing process (Bhattacharya,
(1) Srivastava, and Prasad, 2012; Jin and Jin, 2016b). The
relationship between software projects and software
where, β is fixed quotient of the software project, LOC testing can be understood in the following ways:
is total number of lines of codes to execute “n” num-
ber of functions under fρ number of function points. • Quality assurance: Software testing is essential for
ensuring the quality of the software product be-
• Fuzzified OOPS metrics: Object-oriented pro- ing developed within a software project. It helps
gramming systems (OOPS) metrics refer to identify defects, errors, and vulnerabilities in the
measurements used to evaluate the quality and software, allowing developers to address these is-
complexity of object-oriented software. Fuzzi- sues before the software is released to users
fied OOPS metrics involve applying fuzzy logic • Verification and validation: Software testing is
to these metrics to handle imprecise or uncertain a means of verifying that the software is being
data. Fuzzy logic allows for handling vagueness developed correctly (verification) and validating
in software quality attributes by assigning degrees that it meets the user’s needs (validation). It helps
of membership to different categories, providing confirm that the software aligns with the project’s
a more flexible and nuanced understanding of requirements and objectives
software complexity • Risk mitigation: Software projects inherently
• Cosmic function point (CFP)-based factor analy- involve risks, including the risk of defects or er-
sis and selection: Cosmic function points (CFP) rors. Effective testing helps mitigate these risks by
are a variation of traditional function points used catching and addressing issues early in the devel-
to measure the functional size of a software ap- opment process, reducing the chances of critical
plication based on its business functionality. Fac- failures after deployment
tor analysis and selection in the context of CFP • Iterative development: Many modern software
involves identifying and assigning appropriate development methodologies, such as Agile and
complexity factors to account for variations in DevOps, promote iterative and incremental de-
software projects. These factors help in adjust- velopment. Testing is performed throughout these
ing the functional size measurement to reflect the iterations to continuously assess the software’s
software’s unique characteristics progress and maintain its quality
• COCOMO analysis and new OOPS metrics: • Documentation: Software testing generates docu-
COCOMO (constructive cost model) is a soft- mentation about the software’s behavior, test cas-
ware cost estimation model used to predict the es, and results. This documentation is valuable for
effort, cost, and schedule required for software project managers, developers, and stakeholders to
development. It considers various factors like the track progress and make informed decisions
size of the project, development team experience, • Resource allocation: Software projects need to
and complexity. In the context of object-oriented allocate resources, including time and effort, for
programming, COCOMO can be used to esti- testing activities. The scope and depth of testing
mate effort based on new OOPS metrics, which depend on the project’s requirements and priori-
are measurements specific to object-oriented soft- ties
ware. These new metrics might include measures • Feedback loop: Testing provides feedback to the
of class complexity, coupling, cohesion, and other development team about the software’s perfor-
object-oriented design attributes. mance, functionality, and usability. This feedback
loop helps developers improve the software and
enhance user satisfaction
III. Software testing effort
• The discussion shows that the software testing is
Software testing effort refers to the resources, time, a critical aspect of software projects that ensures
and activities required to effectively test a software the quality, reliability, and functionality of the
application or system to ensure its quality, function- software being developed. It supports the overall
ality, and reliability (Nassif et al., 2019; Cibir and success of the project by identifying and address-
Ayyildiz, 2022). It’s an essential phase of the soft- ing issues, mitigating risks, and providing valu-
ware development life cycle that aims to identify able insights for continuous improvement.
86 Selection of effective parameters for optimizing software testing effort estimation
Table 12.1 Comparative analysis of existing studies
Badri, M., Toure, F., Predict unit testing Regression analysis Predictive model Limited sample size
and Lamontagne, L. effort levels of classes for testing effort
estimation
Bhattacharya, P., Estimate software test PSO (Particle swarm PSO-based estimation Requires tuning PSO
Srivastava, P. R., and effort optimization) of test effort parameters
Prasad, B.
Bluemke, I. and Review and summarize Survey and review Overview and Lack of original
Malanowska, A. testing effort categorization of research data
estimation techniques
Borade, J. G. and Provide an overview of Review of Overview of software Limited focus on
Khalkar, V. R. effort estimation estimation effort estimation specific estimation
techniques methods techniques
Liao, X. and Naseem, Review COCOMO Review of Overview of Limited focus on
A. models and extensions COCOMO models COCOMO models COCOMO models
and extensions
Satapathy, S. M., Early-stage software Random forest Early-stage effort Limited to use
Acharya, B. P., and effort estimation estimation with use case point-based
Rath, S. K. case points estimation
Sharma, A. and Develop metric suite Requirement Metric suite for early Limited validation of
Kushwaha, D. S. for testing estimation engineering estimation of testing the metric suite
document
Singh, V., Kumar, V., Select influential Fuzzy logic and Parameter selection for Limited to parameter
and Singh, V. B. testing parameters AHP-TOPSIS influencing testing selection
Srivastava, P. R., Estimate test effort Bat algorithm Test effort estimation Limited to bat
Bidwai, A., Khan, A., using bat algorithm based on bat algorithm algorithm
Rathore, K., Sharma,
R., and Yang, X. S.
Integration complexity When the systems are integrated with STE is governed by the extent
Software
effort
Automation Test automation initially require higher STE gradually decreases with
testing effort, however, eventually, the the passage of time
manual software testing effort gets
reduced with the passage of time
Skilled workforce The skilled testers may reduce the STE is directly proportional to
testing phase that otherwise may get the length of testing phase
prolonged when it lacks in skilled
testers
Time constraints Tight software project schedule limits STE is inversely proportional
the testing window and increases the to time constraint
software testing effort
Test data Availability of diverse test data is STE is inversely proportional
5 essential. The generation of test data to the availability of test data
needs additional effort that increase
the overall testing effort. Here, data
Resource Availability
Testing configuration Hard and complex software set-up STE is governed by the
configuration can increase software software set-up environment
testing effort during configuration and configuration
Testing tools Availability of suitable tools such as STE is governed by the
hardware, software network resources suitability of testing tools
Testing infrastructure
Multipliers: Developer skill, non-functional needs, the adjustments needed based on factors like ASLOC
platform familiarity, and other factors are reflected in and AT, as explained earlier. The COCOMO II model
multipliers. primarily relies on your estimation of the software
project’s size, measured in thousands of Source Lines
• RCPX – product reliability and complexity; of Code (KSLOC), to calculate the required effort in
• RUSE – the reuse required; terms of Person–Months (PM). In essence, the model’s
• PDIF – platform difficulty; effort estimation heavily depends on your assessment
• PREX – personnel experience; of the project’s scale, as quantified by the size of the
• PERS – personnel capability; codebase.
• SCED – required schedule;
• FCIL – the team support facilities. (5)
2) The reuse model Eaf stands for “Effort Adjustment Factor,” which
The reuse model is an approach in software devel- is derived from the cost drivers. The exponent E in
opment that focuses on leveraging existing software the formula is determined by the five scale drivers.
components, modules, or solutions to enhance effi- Estimation techniques, such as expert judgment, his-
ciency, reduce development time, and improve overall torical data analysis, and specialized software tools,
software quality. It centers on the idea that by reusing can aid in determining the effort required for the
well-tested and proven components, developers can reuse model. In essence, the reuse model’s estimation
avoid reinventing the wheel and instead build upon involves evaluating the integration effort, customiza-
established solutions. The reuse model encourages the tion needs, and associated activities when incorporat-
systematic identification, selection, and integration of ing existing components into a new project.
reusable assets to streamline the development process. With a holistic approach, taking into account the
intricacies of the software, the testing strategy, the
a. Reuse model estimation: team’s capabilities, and the project context, the paper
Estimating effort and resources for the reuse model contributes to provide foundation for the accurate
involves considering factors unique to integrating and testing effort estimation. Regular review and adjust-
adapting reusable components. This estimation model ment of estimates based on evolving project dynamics
uses Equation (3) to estimate the effort. further contribute to improved project planning and
successful software delivery.
(3) V. Conclusion
The paper presents a detailed analysis of various fac-
ASLOC stands for “Actual Source Lines of Code,” tors that are critical for accurate estimation of soft-
which represents the total number of lines of ware testing effort that plays a critical role in project
code that have been created for a software proj- planning and management. The paper delved into
ect. AT refers to the “Proportion of Automatically existing research presented by the research commu-
Generated Code,” indicating the fraction of code that nity in predicting software testing effort estimation.
is generated through automated tools or processes. The paper claims important contribution in laying
ATPROD represents “Engineers’ Productivity in Code down the foundation and preliminary factor analysis
Integration,” signifying how efficiently developers prior to estimating the testing effort for a software
integrate this code. If the estimation process is based project. The multifaceted nature of software projects
solely on manually written code, the estimated Lines necessitates a comprehensive approach to estimation,
of Code (LOC) can be determined using Equation (4). encompassing parameters such as project complexity,
size, requirements volatility, team expertise, and his-
torical data analysis. By considering these parameters,
(4)
organizations can enhance their ability to create more
reliable and realistic testing effort estimates. As the
ESLOC stands for “Estimated Source Lines of Code,” software development landscape continues to evolve,
which refers to the calculated number of lines of code the parameters influencing testing effort estimation
expected in a software project. The costs associated are subject to change with the passage of time and
with modifying reused code, understanding how to advancement of technology. Therefore, a proactive
integrate it, and making decisions about its reuse stance towards continuous improvement and adap-
are considered in determining the adaption adjust- tation is necessary. The collaboration between devel-
ment multiplier. This multiplier takes into account opment and testing teams, ongoing communication,
90 Selection of effective parameters for optimizing software testing effort estimation
and learning from each estimation cycle’s outcomes review. J. Comput., 3(5), 683–693. Available: http://
are crucial for refining the estimation process over [Link].
time. Altogether, a successful software testing effort Mensah, Solomon, Jacky Keung, Kwabena Ebo Bennin, and
estimation demands a harmonious blend of empiri- Michael Franklin Bosu. (2016). Multi-objective opti-
mization for software testing effort estimation. SEKE,
cal analysis, domain expertise, and technological
1–6. doi: 10.18293/SEKE2016-017
advancements. With fine-tuning of the parameters
Brar, P. S., Shah, B., Singh, J., Ali, F., and Kwak, D. (2022).
discussed in this paper, organizations can pave the Using modified technology acceptance model to eval-
way for more accurate STE estimations, leading to uate the adoption of a proposed IoT-based indoor
better resource allocation, project planning, and ulti- disaster management software tool by rescue work-
mately, the delivery of high-quality software systems. ers. Sensors, 22(5), 1866. [Link]
In future, the study can be followed for evaluating the s22051866.
effect of integrating concept of machine learning in Mohammed, N. M., Niazi, M., Alshayeb, M., and Mah-
STE estimation works. mood, S. (2017). Exploring software security ap-
proaches in software development lifecycle: A sys-
tematic mapping study. Comp. Stand. Interf., 50,
References 107–115. doi: 10.1016/[Link].2016.10.001.
Badri, M., Toure, F., and Lamontagne, L. (2015). Predict- Nassif, A. B., Azzeh, M., Idri, A., and Abran, A. (2019). Soft-
ing unit testing effort levels of classes: An exploratory ware development effort estimation using regression
study based on multinomial logistic regression model- fuzzy models. Computat. Intel. Neurosci., 2019.
ing. Proc. Comp. Sci., 62, 529–538. Rajamanickam, Leelavathi. (2016). Principles and Goals of
Bhattacharya, P., Srivastava, P. R., and Prasad, B. (2012). Software Testing. International Journal of Advanced
Software test effort estimation using particle swarm Engineering, Management and Science, 2(5): 239455.
optimization. Adv. Intel. Soft Comput., 132, 827–835. 427–430.
doi: 10.1007/978-3-642-27443-5_95/COVER. Satapathy, S. M., Acharya, B. P., and Rath, S. K. (2016). Ear-
Bluemke, I. and Malanowska, A. (2021). Software testing ly stage software effort estimation using random for-
effort estimation and related problems. ACM Comput. est technique based on use case points. IET Software,
Sur. (CSUR), 54(3). doi: 10.1145/3442694. 10(1), 10–17. doi: 10.1049/IET-SEN.2014.0122.
Borade, Jyoti G., and Vikas R. Khalkar. (2013). Software Sharma, A. and Kushwaha, D. S. (2011). A metric suite for
project effort and cost estimation techniques. Interna- early estimation of software testing effort using re-
tional Journal of Advanced Research in Computer Sci- quirement engineering document and its validation.
ence and Software Engineering, 3(8), 730–739. 2011 2nd Int. Conf. Comp. Comm. Technol., 373–
Cibir, E. and Ayyildiz, T. E. (2022). An empirical study on 378. doi: 10.1109/ICCCT.2011.6075150.
software test effort estimation for defense projects. Singh, V., Kumar, V., and Singh, V. B. (2023). A hybrid novel
IEEE Acc., 10, 48082–48087. doi: 10.1109/AC- fuzzy AHP-TOPSIS technique for selecting parame-
CESS.2022.3172326. ter-influencing testing in software development. Dec.
Singh, J., Goyal, G., and Gill, R. (2020). Use of neuro- Anal. J., 6, 100159. doi: 10.1016/[Link].2022.
metrics to choose optimal advertisement method for 100159.
omnichannel business. Enterp. Inform. Sys., 14(2), Srivastava, P. R., Bidwai, A., Khan, A., Rathore, K., Shar-
243–265. [Link] ma, R., and Yang, X. S. (2014). An empirical study
40392. of test effort estimation based on bat algorithm.
Hidmi, O. and Sakar, B. E. (2017). Software development ef- Int. J. Bio-Ins. Comput., 6(1), 57–70. doi: 10.1504/
fort estimation using ensemble machine learning. Int. IJBIC.2014.059966.
J. Comput. Commun. Instrum. Engg., 4(1), 143–147. Suri, R., Pushpa, and Harsha, S. (2015). Object oriented
Jin, C. and Jin, S. W. (2016). Parameter optimization of software testability (OOSTE) metrics analysis. Int. J.
software reliability growth model with S-shaped test- Comput. Appl. Technol., 4(5), 359–367.
ing-effort function using improved swarm intelligent Thakore, D. and Upadhyay, A. R. (2013). A framework to
optimization. Appl. Soft Comput., 40, 283–291. doi: analyze object-oriented software and quality assur-
10.1016/[Link].2015.11.041. ance. Int. J. Inn. Technol. Explor. Engg. (IJITEE), 5,
Jin, C. and Jin, S.-W. (2016). Parameter optimization of 254–258.
software reliability growth model with S-shaped test- Trendowicz, J., Münch, J., and Jeffery, R. (2011). State of
ing-effort function using improved swarm intelligent the practice in software effort estimation: A survey
optimization. Appl. Soft Comput., 40, 283–291. and literature review. Lecture Notes – Comp. Sci.
Kaner, C., Falk, J., an Nguyen, H. Q. (1999). Testing com- (including subseries Lecture Notes in Artificial Intel-
puter software. John Wiley & Sons. ligence and Lecture Notes in Bioinformatics), 4980
Liao, X. and Naseem, A. (2012). Software models, exten- LNCS, 232–245. doi: 10.1007/978-3-642-22386-
sions and independent models in cocomo suite: A 0_18/COVER.
13 Automated detection of conjunctivitis using convolutional
neural network
Rajesh K. Bawa1 and Apeksha Koul2,a
1
Department of Computer Science, Punjabi University, Patiala, Punjab, India
2
Department of Computer Science and Engineering, Punjabi University, Patiala, Punjab, India
Abstract
Conjunctivitis, commonly referred to as “pink eye,” is a prevalent and contagious eye condition that affects millions world-
wide. Detecting conjunctivitis early and accurately is vital for timely intervention and effective management. Here, a convo-
lutional neural network (CNN) model has been customized to automate the detection of conjunctivitis using eye images. Our
dataset encompasses a diverse array of eye images, which include both healthy and conjunctivitis-affected cases. To tackle
the challenge of limited data, we employ data augmentation techniques to expand the dataset. After pre-processing and aug-
mentation, we curate a collection of 5135 eye images representing both pink-eye pathology and healthy states. Subsequently,
these augmented images undergo classification using the developed CNN model. During execution, the customized CNN
model obtains an impressive accuracy of 88.80%, with a loss of 0.25, and demonstrates precision, recall, and F1 scores of
0.50. The CNN model holds promise as an automated solution for conjunctivitis detection. Its accuracy and efficiency could
substantially support medical professionals in early diagnoses, facilitating timely treatment and curbing transmission rates.
a
apekshakoulo9@[Link]
92 Automated detection of conjunctivitis using convolutional neural network
B. Data pre-processing
The eye data collected are of different sizes which can
hamper the performance of the system, hence their
size have been reduced to (224×224), as shown in
Figure 13.5.
Later, the quality of the resized images have been
enhanced and the process starts with individual histo-
gram equalization on each color channel (blue, green,
and red) to enhance contrast by spreading pixel inten-
Figure 13.4 Number of images in the dataset sity distribution. This enriches visual appeal, making
dark and light areas distinct. Following this, unsharp
masking is applied. A blurred version of the enhanced
A. Dataset image is subtracted from the original, emphasizing
The dataset employed in this research has been care- high-frequency components like edges and details.
fully customized by gathering specific images that These components are then blended back into the
show pink eye from the eye diseases virus dataset vol- image, sharpening it and highlighting features. This
ume 1 (Kaggle, 2020). Moreover, a separate collection blend of techniques enhances contrast and sharpness,
of images depicting healthy eyes has been acquired resulting in an image with heightened visual appeal
from various reputable online sources, as shown in and clear details. The approach is demonstrated by
Figure 13.3. showcasing the original and enhanced images, as
A total of 356 .jiff images illustrating instances of shown in Figure 13.6.
pink eye were amassed for analysis. It is important to
mention that a small subset of these images was con- C. Data augmentation
sidered unsuitable and was manually excluded from It involves applying transformations to the original
the dataset, which ends up with the 265 number of images to diversify the dataset and improve model
images. Furthermore, an independent set of 130 .jpg training. Here, ImageDataGenerator() has been used
images featuring healthy eyes (Singh et al., 2019) was to perform rotation (within a range of -50 to +50
obtained for the purpose of comparison, as shown in degrees), horizontal flipping, and vertical flipping
Figure 13.4. generating 12 augmented images of single image (4
94 Automated detection of conjunctivitis using convolutional neural network
Table 13.1 Layered architecture of proposed CNN model
Metrics Formulae
Accuracy
Loss
Precision
Recall
F1 score
Hyper-parameters Values
Activation ReLu/Sigmoid
Optimizer Adam
Class mode Binary
Figure 13.9 Learning curves of CNN model
Batch_Size 32
Loss Binary cross entropy
Dropout rate 0.5 Table 13.5 Performance summary.
Epochs 10 Model Recall Precision F1 score
Similarly, the model has been also evaluated for consideration. The model’s performance may vary
binary class of the dataset i.e., for healthy eye and with factors like dataset size, diversity, and quality.
pink eye on the basis of precision, recall, and F1 score Over fitting remains a concern, especially if the data-
in Table 13.6. set is small or imbalanced.
Abstract
A wireless sensor network (WSN) is a key technology in the implementation of several applications, including light-duty data
streaming applications and straightforward event/phenomena monitoring systems. The energy-efficiency of wireless sensor
networks is a crucial issue in case of development and implementation. The goal of this effort is to increase the information
processing and routing process of energy efficiency. This research paper’s primary goal is to provide a complete review of
WSN. This article gives a broad overview of the WSN and some of its key features. This study also discusses several WSN
threats, WSN research obstacles, and WSN applications.
[Link]@[Link]
a
Applied Data Science and Smart Systems 99
a) Processing unit
This part often acts as the sensor node’s heart-
beat. Its internal microcontroller processes the
data that is sent to it. One or more of the mi-
crocontrollers that are most often used in sensor
nodes include the MSP 430, Intel Strong ARM,
and SA-1100. To store the instruction set, a flash
memory is also linked to this device. Figure 14.4 Applications areas of WSN
Applied Data Science and Smart Systems 103
for purposes like weather observation. If the price X. Mobile sinks in WSN: challenges
was reasonable overall for sensor networks, it would
Location Identification: To transmit the detected
be much more reasonable and successful for consum-
data, the sensor nodes require the availability of the
ers who demand careful consideration.
mobile actuator. The broadcast techniques are used
by mobile sinks to relay their position to network
Placement
nodes. Unfortunately, these methods need a signifi-
Node location in WSN could be a simple problem to
cant amount of resources to broadcast position data
tackle. Special strategies are required to position and
from the sink to the network on a regular basis. The
handle a broad spectrum of nodes in a relatively lim-
network requires a lot of message forwarding from
ited environment. In a highly sensitized region, 100 to
each node, which makes it difficult to use resources
10,000 of sensors have also been placed.
efficiently. By using an overhearing method, it takes
less broadcast messages to locate a mobile sink. The
Restrictions on strategy
mobile sink creates beacons with fresh position data
The creation of smaller, more affordable, and more
and transmits them to nearby access points. The access
effective devices is the main goal of wireless sensor
points that are in the mobile sink’s communication
design. The design of WSN will be influenced by a vari-
range receive the beacon and alter their message head-
ety of additional competitions. WSN has encountered
ers to point at the mobile sink. The remaining nodes
limited-restriction hardware and software approach
in the networks locate the mobile sink’s new position
paradigms (Tyagi and Kumar, 2013).
after hearing the changed header. Fewer fixed points
are used in the footprint-based technique. The sensor
IX. Clustering architecture in WSN nodes use fixed positions to determine the communi-
Clustering is the practice of assembling sensor nodes cation channel to the mobile sink. Nodes determine
that are geographically adjacent to one another their logical coordinates and the path to the mobile
(Gao et al., 2010; Gardašević et al., 2020), (Surya sink based on the fixed locations. The overhead
Engineering College & Institute of Electrical and caused by the location identification protocol should
Electronics Engineers, n.d.). A cluster head (CH) is a be maintained and considered to reduce the need for
node that manages a cluster and may start the cluster- retransmission of broadcast messages and extend net-
ing process. Cluster members are the residual nodes in work lifespan.
the cluster. These CM nodes will continually perceive Routing and mobile sink trajectory: The mobile sink
their surroundings and transmit data to the corre- trajectory is essential for the routing of sensed data.
sponding CH nodes. As all member nodes in a cluster In contrast to large-scale WSNs, a mobile sink may
are close to one another, the information they provide travel to each node in the net to collect the sensed
will also be redundant. In the majority of application data. Using special nodes that are fewer than the total
instances, sending this duplicate information to the number of nodes in the network, the random walks-
BS is unnecessary, and it also shortens the network’s based strategy decouples the mobile sink route from
lifespan. In order to create a single piece of infor- the sensor nodes. Movable sinks may roam around
mation, the CH nodes combine the data they have freely and collect data from a desired number of
obtained from the CMs. The BS will only be informed nodes (which act as a sub sinks or access points). By
of this one piece of information. In general, WSNs using position identification methods, sensor nodes
benefit from the clustering architectural paradigm in locate the mobile sink. In order to enhance the dura-
the following ways. tion of WSNs, the trajectory and routing path choices
are crucial. The mobile sink’s trajectory must guaran-
Conserving bandwidth: It occurs when a network is tee sink availability across the network with the least
grouped and nodes are logically segmented and con- amount of routing overhead.
nected to their respective groups. As a result, the clus- Transmission scheduling: The nodes in WSNs transfer
ters may share the same communication bandwidth the detected data to the mobile sink as soon as it enters
without encountering any interference. To prevent their communication range. WSNs are resource-con-
inter-cluster interferences, each cluster will have its strained networks that strive to cut down on energy
own spread code. use while distributing data to a mobile sink. Till the
Scalability: After the first deployment, new sets of mobile sink arrives at the nearest point in the commu-
nodes are included in the network to improve the nication range, the nodes will not be able to send their
accuracy level of the information. The newly added data according to the transmission scheduling scheme.
nodes may be readily accommodated using clustering
strategy as a new cluster or included in the available When more than one node recognizes the
groups during the process. mobile sink within its communication range, data
Applied Data Science and Smart Systems 105
Figure 14.6 Multiple nodes request for transmission Figure 14.8 Mobile sink revoke compromised node
channel at the same time and broadcast control messages to the network
transmission in WSNs with densely placed sensor Security: Because the sink is mobile, WSNs are vulner-
nodes becomes challenging. Data scheduling tech- able to a variety of attacks. For WSNs with mobile
niques addressed how the mobile sink in densely dis- sinks, conventional security techniques are insuf-
tributed WSNs, as depicted in Figure 14.5, chooses ficient (Abella et al., 2019). A mobile sink may be
the transmission channel. The mobile sink assigns a compromised to access the whole network. If WSNs
channel based on the measure of data to be trans- deploy numerous mobile sinks, seizing one will give
mitted and the node’s remaining power. Effective data the attacker access to a significant chunk of the net-
distribution in WSNs requires a distance and speed work. Two distinct key pools are generated by the
traveled – Finding a reasonable compromise between key management method, one for connecting sensor
efficient data aggregation and the network’s capacity nodes and access points and the other for connecting
to access the mobile sink is a difficult task. Mobile mobile sinks and access points. With the sink’s mobil-
sink must remain inside the sensor nodes’ communi- ity comes an increase in security complexity. Figure
cation range until the nodes have finished transmit- 14.5 depicts the case of an access point being taken
ting. The base station must respond quickly to WSN over and a duplicated mobile sink being brought into
applications in order to manage the regions of inter- the network by an attacker. Researchers need to pay
est. The mobile sink moves along the trajectory at a close attention to WSN security.
controlled speed thanks to transmission scheduling Upkeep of the network: The mobile sink’s privilege
and routing algorithms, forcing the sensor nodes to level must be set before to deployment in order to
transfer any observed data to the sink. To extend the allow the base station to manage the network. Figure
lifespan of the network, a trajectory with a minimal 14.6 illustrates how mobile sinks are utilized for node
trip distance and maximum network coverage must revocation as well as broadcasting private control
be chosen. The suggested model enhances the perfor- messages to the whole network during major security
mance in terms of the speed and distance that mobile threats.
sinks move.
106 An overview of wireless sensor networks applications, challenges and security attacks
The need for mobile sinks in WSNs is rapidly grow- XII. Conclusion
ing. The introduction of mobile sink in WSNs pres-
WSN is a key technology in the implementation of
ents further difficulties for networks. The difficulties
several applications, including light-duty data stream-
of implementing mobile sinks in WSNs are noted.
ing applications and straightforward event/phenom-
WSNs still experience resource depletion when using
ena monitoring systems. The energy-efficiency of
the current techniques (Figures 14.7 and 14.8).
wireless sensor networks is a crucial issue while devel-
oping and running them. The goal of this effort is to
XI. General basic features increase the information processing and routing pro-
A. Sensor network architecture cess’ energy efficiency. The primary goal of this work
Any of the following methods are used by sensors to is to provide a complete review on WSN.
send data to the base station.
References
• In a flat adhoc architecture, sensors work togeth-
Abella, C. S., Bonina, S., Cucuccio, A., D’Angelo, S., Gi-
er to send data to the access point, also known as ustolisi, G., Grasso, A. D., Imbruglia, A., Mauro, G.
the base station (AP or BS), and several hops are S., Nastasi, G. A. M., Palumbo, G., Pennisi, S., Sor-
used in the transmission process. bello, G., and Scuderi, A. (2019). Autonomous energy-
• Sensors are grouped into clusters in hierarchi- efficient wireless sensor network platform for home/
cal networks, and cluster heads are in charge office automation. IEEE Sens. J., 19(9), 3501–3512.
of gathering and aggregating data from sensors [Link]
and reporting to the access point. In hierarchi- Ahmad, Arshad, Ayaz Ullah, Chong Feng, Muzammil Khan,
cal networks, single-hop transmission to the BS Shahzad Ashraf, Muhammad Adnan, Shah Nazir, and
and multihop transmission between clusters are Habib Ullah Khan. (2020). Towards an improved
used. energy efficient and end-to-end secure protocol for
iot healthcare applications. Security and Commu-
• In a sensor network with mobile access, mobile
nication Networks. vol. 2020: 1–10. [Link]
BS roaming the sensor field directly communi- org/10.1155/2020/8867792
cates with the sensors, and transmission from the Ahmed, G., Zhao, X., Fareed, M. M. S., Asif, M. R.,
sensors to BS is one-hop. and Raza, S. A. (2019). Data redundancy-control
energy-efficient multi-hop framework for wireless
B. Wireless sensor nodes sensor networks. Wire. Person. Comm., 108(4),
Motes are another name for sensor nodes. The mar- 2559–2583. [Link]
ket offers a wide variety of sensors. The following are 06538-0.
some of the sensor nodes in use right now: Ahmed, M., Salleh, M., and Channa, M. I. (2018). Rout-
ing protocols based on protocol operations for un-
University of California Los Angeles (UCLA) Wireless derwater wireless sensor network: A survey. Egyptian
Integrated Network Sensors (WINS) UCLA created Informat. J., 19(1), 57–62. [Link]
some of the low power wireless integrated micro sen- eij.2017.07.002.
Al Qundus, J., Dabbour, K., Gupta, S., Meissonier, R.,
sors that are utilized as sensor nodes. Using a 1 mW
and Paschke, A. (2022). Wireless sensor network for
transmitter, it provides wireless communication at a AI-based flood disaster detection. Ann. Oper. Res,
speed of 100 Kbps over a distance of 10 m. 319(1), 697–719. [Link]
UC Berkeley quotes: Crossbow’s Mica family con- 020-03754-x.
tains the members Mica, Mica2, Mica2Dot, and Alghamdi, T. A. (2020). Energy efficient protocol in wire-
MicaZ. With a 16 MHz CPU, 4 KB of RAM, 2 KB less sensor network: Optimized cluster head selection
of flash memory, and a data rate of up to 250 Kbps, model. Telecomm. Sys., 74(3), 331–345. [Link]
MicaZ supports the IEEE 802.15.4 standard and org/10.1007/s11235-020-00659-9.
Ali, Ahmad, Yu Ming, Sagnik Chakraborty, and Saima
ZigBee protocols.
Iram. (2017). A comprehensive survey on real-time
AMPS from MIT: It is a low-end standalone guard- applications of WSN. Future internet. 9(4): 77 pp.1-
ing node that can also act as a fully complete node 22. [Link]
for middle-end sensor networks or as a supporting Al-Turjman, F. (2019). Cognitive-node architecture and a de-
element in a higher-end sensor system. The institution ployment strategy for the future WSNs. Mobile Netw.
uses these motes for a variety of purposes. Appl., 24(5), 1663–1681. [Link]
s11036-017-0891-0.
Tiny node: 512 KB of external flash memory, 8 KB Behera, T. M., Mohapatra, S. K., Samal, U. C., Khan, M.
of RAM for programmes and data, and the TinyOS S., Daneshmand, M., and Gandomi, A. H. (2020).
operating system. Moreover, there are BTNode, I-SEP: An improved routing protocol for heteroge-
Imote, Iris Mote, TelosB, Wasp Mote, and others. neous WSN for IoT-based environmental monitoring.
Applied Data Science and Smart Systems 107
IEEE Internet of Things J., 7(1), 710–717. [Link] for precision agriculture: A review. Sensors, 17(8), 1–45:
org/10.1109/JIOT.2019.2940988. 1781. doi: [Link]
Benayache, A., Bilami, A., Barkat, S., Lorenz, P., and Ta- Bajaj, Karan, Bhisham Sharma, and Raman Singh. (2020).
leb, H. (2019). MsM: A microservice middleware for Integration of WSN with IoT applications: a vision,
smart WSN-based IoT application. J. Netw. Comp. architecture, and future challenges. Integration of
Appl., 144, 138–154. [Link] WSN and IoT for Smart Cities. 79–102. doi: https://
jnca.2019.06.015. [Link]/10.1007/978-3-030-38516-3_5
Chaitra, M., and B. Sivakumar. (2017). Disaster debris de- Preeth, SK Sathya Lakshmi, R. Dhanalakshmi, R. Kumar,
tection and management system using wsn & iot.. and P. Mohamed Shakeel. (2018). An adaptive fuzzy
International Journal of Advanced Networking and rule based energy efficient clustering and immune-in-
Applications. 9(1): 3306–3310. spired routing protocol for WSN-assisted IoT system.
Díaz, S. E., Pérez, J. C., Mateos, A. C., Marinescu, M. C., Journal of Ambient Intelligence and Humanized Com-
and Guerra, B. B. (2011). A novel methodology for puting. vol. (2018): 1–13. [Link]
the monitoring of the agricultural production process s12652-018-1154-z
based on wireless sensor networks. Comp. Elec. Ag- Shahraki, A., Taherkordi, A., Haugen, O., and Eliassen, F.
ricul., 76(2), 252–265. [Link] (2021). A survey and future directions on clustering:
pag.2011.02.004. From WSNs to IoT and modern networking para-
Elappila, M., Chinara, S., and Parhi, D. R. (2018). Surviv- digms. IEEE Trans. Netw. Ser. Manag., 18(2), 2242–
able path routing in WSN for IoT applications. Pervas. 2274. [Link]
Mobile Comput., 43, 49–63. [Link] Shukla, A. and Tripathi, S. (2020). A multi-tier based cluster-
pmcj.2017.11.004. ing framework for scalable and energy efficient WSN-
Elappila, M., Chinara, S., and Parhi, D. R. (2020). Surviv- assisted IoT network. Wireless Netw., 26(5), 3471–
ability aware channel allocation in WSN for IoT ap- 3493. [Link]
plications. Pervas. Mobile Comput., 61. [Link] Khan, JavedAkhtar. (2019). —Multiple Cluster-Android
org/10.1016/[Link].2019.101107. lock Patterns (MALPs) for Smart Phone Authentica-
Farhan, L., Shukur, S. T., Alissa, A. E., Alrweg, M., Raza, U., tion. In 2019 3rd International Conference on Com-
and Kharel, R. (2017). A survey on the challenges and puting Methodologies and Communication (ICCMC).
opportunities of the Internet of Things (IoT). Proc. 619–623. IEEE. doi: 10.1109/ICCMC.2019.8819635.
Int. Conf. Sens. Technol., ICST, 2017-December, 1–5. Sutagundar, A. V. and Manvi, S. S. (2013). Wheel based
[Link] event triggered data aggregation and routing in wire-
Gao, Q., Zuo, Y., Zhang, J., and Peng, X. H. (2010). Improv- less sensor networks: Agent based approach. Wire.
ing energy efficiency in a wireless sensor network by Per. Comm., 71(1), 491–517. [Link]
combining cooperative MIMO with data aggregation. s11277-012-0825-x.
IEEE Trans. Vehicul. Technol., 59(8), 3956–3965. Tian, H., Shen, H., and Roughan, M. (2008). Maximizing
[Link] networking lifetime in wireless sensor networks with
Gardašević, G., Katzis, K., Bajić, D., and Berbakov, L. regular topologies. Par. Distribut. Comput. Appl. Tech-
(2020). Emerging wireless sensor networks and in- nol., PDCAT Proc., 211–217. [Link]
ternet of things technologies—foundations of smart PDCAT.2008.29.
healthcare. Sensors (Switzerland), 20(13), 1–30. Tyagi, S. and Kumar, N. (2013). A systematic review on
[Link] clustering and routing techniques based upon LEACH
Haseeb, Khalid, Ikram Ud Din, Ahmad Almogren, and Nav- protocol for wireless sensor networks. J. Netw. Comp.
eed Islam. (2020). An energy efficient and secure IoT- Appl. 36(2), 623–645. [Link]
based WSN framework: An application to smart ag- jnca.2012.12.001.
riculture. Sensors, 20(7): 2081 pp. 1–14. doi: https:// Weng, C. E. and Lai, T. W. (2013). An energy-efficient rout-
[Link]/10.3390/s20072081 ing algorithm based on relative identification and
Azzam, Riad, and Nabil Aouf. (2014). Embeded fusion of direction for wireless sensor networks. Wire. Per.
visual and acoustic for active acoustic source detec- Comm., 69(1), 253–268. [Link]
tion with SGGMM. In Proceedings ELMAR-2014. s11277-012-0571-0.
1–4. IEEE. doi: 10.1109/ELMAR.2014.6923352 Yu, X., Wu, P., Han, W., and Zhang, Z. (2013). A survey
Jaiswal, K. and Anand, V. (2020). EOMR: An energy-effi- on wireless sensor network infrastructure for agricul-
cient optimal multi-path routing protocol to improve ture. Comp. Stan. Interf., 35(1), 59–64. [Link]
QoS in wireless sensor network for IoT applications. org/10.1016/[Link].2012.05.001.
Wire. Pers. Comm., 111(4), 2493–2515. [Link] Zhou, Z., Zhou, S., Cui, S., and Cui, J. H. (2008). Energy-
org/10.1007/s11277-019-07000-x. efficient cooperative communication in a clustered
Singh, J., Singh, S., Singh, S., and Singh, H. (2019). Evalu- wireless sensor network. IEEE Transac. Vehicular
ating the performance of map matching algorithms Technol., 57(6), 3618–3628. [Link]
for navigation systems: an empirical study. Spatial In- TVT.2008.918730.
form. Res., 27, 63–74. Zhu, C., Zheng, C., Shu, L., and Han, G. (2012). A survey
Jawad, Haider Mahmood, Rosdiadee Nordin, Sadik Kamel on coverage and connectivity issues in wireless sen-
Gharghan, Aqeel Mahmood Jawad, and Mahamod Is- sor networks. J. Netw. Comp. Appl., 35(2), 619–632.
mail. (2017). Energy-efficient wireless sensor networks [Link]
15 Internet of health things-enabled monitoring of vital signs
in hospitals of the future
Amit Sundas1,a, Sumit Badotra2, Gurpreet Singh3 and Amit Verma4
1,3
Department of Computer Science and Engineering, Lovely Professional University, Phagwara, Punjab, India
2
School of Computer Engineering and Technology, Bennett University, Greater Noida, India
Department of Computer Science and Engineering, University Center for Research and Development, Chandigarh
4
Abstract
Vital signs and other extensive patient data are among the many types of information typically obtained by hand in hospitals
utilizing discrete medical equipment. It might be challenging for careers to integrate and analyze this information since it is
often kept in separate spreadsheets and not part of patients’ electronic health records. Connecting medical equipment via a
decentralized network such as the Internet is one way to get around these restrictions. By combining data from many sources,
we can more accurately assess a patient’s health and plan for preventative measures. In this study, we present the notion of
the internet of health things (IoHT) and conduct a broad landscape analysis of the methods that may be used to collect and
integrate data on vital signs in healthcare facilities. The potential use of intelligent algorithms is investigated, and common
heuristic techniques such weighted early warning score systems are addressed. In order to maximize efficiency, make the most
of available resources, and prevent unnecessary patient health decline, this article suggests potential avenues for merging pa-
tient data on hospital wards. It is stated that the IoHT paradigm will continue to provide better options for patient treatment
on hospital wards, and that a patient-centered approach is crucial.
Keywords: Machine learning, sepsis, vital sign, prediction, electronic health records
amitsundas1992@[Link]
a
Applied Data Science and Smart Systems 109
II. Hospital treatment focused on the individual rules determine which vital signs are measured, which
patient is contentious.
Elliott and Coventry recommended eight vital signs
Patient-centered care (PCC) stands as a pivotal
(Rana et al., 2018), adding pain, consciousness, and
hospital quality indicator (Gultepe et al., 2013). The
urine output (Iyer et al., 2022). Patients lived longer
assessment of PCC hinges on factors such as patient
Table 15.1 defines, normalizes, and impacts eight vital
needs, effective provider communication, and the
indicators hospitals may monitor for patient health.
availability of services. From an information tech-
Monitoring vital signs raises challenges about how
nology (IT) perspective, PCC can be equated to the
frequently and what to report. Unlike Table 15.1, not
patient’s EHR. This differs significantly from vari-
all vitals may be obtained instantly.
ous enterprise resource planning (ERP) systems of
Pain assessment is subjective. WILDA verifies pain
the past, which primarily aimed to optimize work-
terms, intensity on a 0–10 scale, location, duration,
flow and procedural aspects (Bloch et al., 2019). In
aggravating factors, and pain-relieving variables
the context of PCC, hospital ward vital signs play a
(Sundas et al., 2021). The patient-caregiver WILDA
crucial role, serving as essential markers for identify-
method may use computerized data recording. Tablets
ing patient health concerns and their correlation with
and smartphones can capture patient data for EHRs.
other pertinent data.
Assessing consciousness requires patient-provider
communication. The Glasgow Coma scale measures
A. Vital signs monitoring
eye opening, verbal, and motor responses (Khan et
Patient-centered care improves hospital treatment
al., 2021). These assessments’ numerical outcomes
(Tang et al., 2020). Patient-centered care is measured
depend on patient reactions to stimuli.
by how well it meets patient needs, how fast clinicians
Electronically capturing these numbers may assist
communicate health data, and how readily patients
evaluate the patient’s neurological condition. Catheters
may get treatments. The EHRs are PCC medical
may automatically record urine output. Manual uri-
information systems. Thus, ERPs enhance workflow
nometers are still used (Sundas et al., 2021).
(Khan et al., 2014).
Hospital ward vital signs, vital signs used to high- B. Patient risk assessment
light patient health issues, and vital signs connected Hospital PCC examines vital signs regularly and more
to other data may describe PCC. Hospital nurses have often if concerns arise. Data and graded response tech-
measured the same vitals since 1900 (Sundas et al., niques lower risk. Monitoring and triggering define
2021). how frequently, what, and when to check in. Table
Blood pressure, temperature, heart rate, and respi- 15.2 outlines healthcare institution risk assessment
ration have oxygen saturation. To effectively assess a approaches. Check metric first (Liu et al., 2014). This
patient’s state, National Institute for Health and Care group uses MET (Medical Emergency Team) calling
Excellence (NICE) suggests monitoring oxygen satu- criteria. MET is determined by airway threats, respi-
ration in addition to the five vitals. Consider urine ratory or cardiac arrest, state alterations, and convul-
output, discomfort, and biochemical testing. Hospital sions (Singh et al., 2019; Sundas et al., 2021). Group
A pain scale is used to measure Pain of level The view from the patient The patient reports no pain
the patients’ pain intensity on the 0–10 pain scale (1–3
mild, 4–6 moderate, 7–10
severe)
The force multiplied by the Blood pressure Variables such as age, posture, 90/60–120/80 mmHg
period between heartbeats effort, sleep, slant, and
(systole) that blood exerts on confounding variables (such as
arteries White-Coat-Syndrome or anxiety)
How many times in 60 seconds Breathing rate Variables such as age, oxygen Breathing rate: 12–1/8/min
the chest moves up and down levels in the environment, pain and
anxiety levels, and physical effort
Estimates the quantity of oxygen SPO2 Workload, oxygen levels, and From 95% to 100%
in the blood by measuring the other confounding variables (such
saturation level of its peripheral as activity and pain intensity)
capillaries (SpO2)
110 Internet of health things-enabled monitoring of vital signs in hospitals of the future
Table 15.2 Methods, frameworks, and systems for assessing risk
EWS + MET, MEWS + PART intensity Combination Monitoring development, graded response, varying
sensitivity and specificity
Acceptable calling standards Single parameter Easy to use, but no improvement tracking
PART Multiple parameter high sensitivity and low specificity, yet it allows
progress tracking and progressive reaction
The worthing physiological scoring Aggregate scoring Allows development monitoring, progressive response,
system (EWS, MEWS, ViEWS) and high sensitivity and specificity, based on the score
2 needs one abnormal vital sign. PART (Prehospital Heuristics underpin all these methods. These meth-
Acuity Rating for Triage) calling requirements are an ods compare physiological measures to preset criteria,
example. The PART method measures respiration, resulting in many false positives. These examinations
heart rate, systolic blood pressure, consciousness, may employ artificial intelligence (AI) to better diag-
oxygen saturation, and urine output. nose the patient.
The third category of risk assessment focuses on
evaluating vital signs to detect early signs of health C. Health information recording
deterioration. The early warning score (EWS) was ini- The PCC programmers include physiological obser-
tially introduced as the scoring system. Notably, EWS vations at admission and throughout hospitaliza-
has gained widespread adoption in most UK hospitals tion. For this reason, healthcare workers employ
due to its endorsement by the NICE, and its proven EHR systems. Healthcare practitioners manage most
effectiveness (Khan et al., 2021). The EWS relies on EHRs. These systems solely monitor the patient’s cur-
calculated data, assessing a patient’s health by ana- rent healthcare provider. Combining data from sev-
lyzing multiple parameters at varying intervals. The eral EHR systems doesn’t provide a comprehensive
vital sign data needed for EWS can be entered into the patient EHR (Chen et al., 2016).
patient’s EHR either manually or automatically, and Personal health records (PHR) are one option for
this allows for continuous calculation and visualiza- achieving a holistic and unified picture of a patient’s
tion of the EWS score over time. The process of man- health. A PHR is an individual’s own representation
aging this data can be carried out using mechanical or of their health records, which may consist of separate
manual EWS devices, including paper-based systems. pieces of data or include data from several other sources.
The initial EWS score is computed based on vital Patients have full authority over their PHRs, including
signs such as systolic blood pressure, temperature, the ability to appoint a proxy or set access privileges.
heart rate, breathing rate, and level of awareness (Liu Involving patients in the management of their health
et al., 2014). Each vital sign is compared to estab- information improves collaboration and participation
lished norms to determine an individual score, with a in therapy (Sundas et al., 2021). Furthermore, the intro-
range of 0–6 for systolic blood pressure and 0–3 for duction of mobile devices and wearables drastically
the remaining parameters. The overall EWS score is alters the patient’s role in engaging with their PHR by
derived by summing up the scores for each vital sign enabling real-time monitoring of vital signs, supple-
and adding them to the respective norms. It’s worth menting health information, and allowing for more
noting that there exist several versions of the EWS. proactive intervention. The current tendency is for
One such variant is the modified early warning score patients to supplement their healthcare providers’ EHR
(MEWS), which incorporates urine output as the sixth data with data collected from their own wearables and
vital indicator. Additionally, the vitalpac early warning mobile devices (Sundas et al., 2022) (Figure 15.1).
score (ViEWS) offers a solution for bedside vital sign
monitoring using a PDA. Another system, the worth- III. Internet of health things
ing physiological scoring system, takes into account a
broader range of vital signs, including respiration rate, Since 1999, the IoT has grown into a worldwide
heart rate, arterial pressure, body temperature, oxy- sensor, wireless communication, and information
gen saturation, and level of awareness, in estimating a processing network. The IoT relies on smart items,
patient’s risk of adverse outcomes. These risk assess- which can transmit and analyze information to
ment methods in the third category combine the ease interact autonomously. Recent efforts to define, IoT
of use from the previous categories with the enhanced have included sensing environmental data, providing
sensitivity provided by the EWS and its variants. communication services, analytics, applications, and
Applied Data Science and Smart Systems 111
NFC, on the other hand, is a contactless proximity possible remedies. The increased complexity and vari-
communication technology that operates at close dis- ety of data has led to a rise in research on big data
tances, typically within approximately 4 cm in prac- analysis in healthcare. We focus on ML approaches
tice (though theoretically, it can work at distances of for data modeling. These approaches typically include
less than 10 cm). The proximity nature of NFC, cou- three steps: data collection, feature selection and
pled with its user-friendliness, makes it an excellent extraction, and learning (Li-wei et al., 2014).
choice for enabling communication between patients IoHT devices and other healthcare equipment with
and their medical data. registration and synchronization capabilities gather
Finally, low-power area networks like 6LoWPAN data. After pre-processing (filtering, standardizing, and
enable the delivery of IPv6 packets in wireless sensor aggregating), features (signal descriptive statistics, tem-
networks (WSNs), extending the reach of the IoT to poral and frequency domain characteristics) that dis-
the level of individual sensor nodes. IPv6’s scalability, criminate the patient’s health condition are identified
improved mobility features, and support for multiple and chosen. A classifier or regressor algorithm is taught
stakeholders’ management have enhanced the admin- to relate the data to health deterioration. Deep learn-
istration of smart objects. Consequently, 6LoWPAN ing uses algorithms to extract information from raw
is widely recognized as the foundational technol- data instead of hand-crafted qualities. After training,
ogy for the IoHT in a substantial body of literature the model may be used as a decision support system
(Baidillah et al., 2023). to evaluate a patient’s health or offer relevant actions.
B. IoHT with real-time health status tracking IV. Algorithms with intelligence for monitoring
The IoHT may help hospitals manage PCC and patient vital signs
data. Nurses manually take vital signs. Manual sphyg-
momanometers, stethoscopes, pain and consciousness Many ML algorithms incorporate critical factors to
questionnaires are employed. In these cases, a smart- enhance their predictive capabilities. Support vector
phone or tablet might help the caretaker by providing machines (SVMs), for instance, are capable of assess-
additional information or collecting data. Electronic ing patient risk by considering various indicators,
vital sign registration saves time and labor. including patient demographics, laboratory findings,
A WSN of wireless personal devices may collect and vital signs. SVMs can predict daily risk ratings for
vital indicators. IPv6 over 6LoWPAN is replacing patients and then aggregate these ratings to stratify
manufacturer specifications for linking smart health overall risk.
devices. Sundas et al., presented a multisensor pain In the healthcare domain, decision trees are com-
assessment approach. These sensors include acceler- monly used for disease classification, enabling the
ometers and GPS trackers for activity levels, micro- identification and categorization of different medi-
phones, and a computer’s capacity to interpret speech cal conditions. Medical research has increasingly
and facial expressions. The authors noted the abun- explored the use of individual neural networks (NN)
dance of smartphone and tablet pain measuring appli- and their integration with other approaches. Notably,
cations. Similar to Aung and colleagues’ technique, one of the early experiments applied long short-term
one may employ image processing to analyze eye memory (LSTM) recurrent neural networks (RNN)
movement, speech analysis to evaluate verbal replies, to detect patterns in EHR data. These NN are well-
and an accelerometer or gyroscope to evaluate motor suited for handling time series data, irregular sam-
responses to determine the patient’s awareness. BP, pling, and gaps in medical records.
temperature, HR, RR, and SpO2 may be monitored To address issues related to missing data, research-
using IoHT. Otero and colleagues recommended auto- ers have developed deep models incorporating gated
matic urine monitoring for severely unwell patients. recurrent units (GRU), which provide effective rep-
RFID, NFC, and Bluetooth can communicate sensor resentations for incomplete or intermittent data. The
data to smartphones. Smartphones convey this data application of RNN to medical data represents a bur-
to a fog or cloud middleware. Intelligent algorithms geoning area of research that continues to evolve and
for processing huge patient data requires further development (Li-wei et al., 2014).
The IoHT, high-throughput sequencing platforms
(genomics, proteomics, and metabolomics), real-time A. Deep learning techniques [11–15] are another op-
imaging, and point-of-care diagnostic devices have tion to consider
made health informatics a data-rich field. Environment NN with several nested layers and neurons lies at the
and social media may provide health information (Liu heart of these approaches (Iyer et al., 2022). Using a
et al., 2014). “Big data analysis” is the processing of large number of neurons enables the coverage of a great
massive amounts of vital sign data using sophisticated deal of raw data, and the option of cascading many
algorithms to identify health decline risks and predict layers enables the automated abstraction of a higher
Applied Data Science and Smart Systems 113
level, eliminating the need for human involvement patients with closed-loop alarm. IoT-enabled Smart
When applied to vital signs data, this function may help Healthcare Sys. Serv. Appl., 143–176.
extract potentially nuanced and obscure insights from Vistisen, S. T., Johnson, A. E. W., and Scheeren, T. W. L.
simple observation. Ravi and coworkers point to con- (2019). Predicting vital sign deterioration with artifi-
cial intelligence or machine learning. J. Clin. Monit.
volutional neural networks (CNNs) as the kind of deep
Comput., 33(6), 949–951.
learning having the most influence on health informatics
Singh, Jaiteg, and Nandini Modi. (2019). Use of information
at the moment (Bloch et al., 2019). CNN has been stud- modelling techniques to understand research trends in
ied extensively, but most of the time it’s used to analyze eye gaze estimation methods: An automated review. He-
photos of the human body for diagnosis. liyon, 5(12), 1–12.
Lujie, C., Dubrawski, A., Wang, D., Fiterau, M., Guillame-
V. Conclusion Bert, M., Bose, E., Kaynar, A. M. et al. (2016). Using
supervised machine learning to classify real alerts and
This review included vital sign monitoring and analy- artifact in online multi-signal vital sign monitoring
sis to the IoHT to predict patient health risks. The data. Crit. Care Med., 44(7), e456.
first portion of the review included the eight main Bloch, Eli, Tammy Rotem, Jonathan Cohen, Pierre Sing-
physiological observations: blood pressure, body tem- er, and Yehudit Aperstein. (2019). Machine learn-
perature, heart rate, respiration rate, oxygen satura- ing models for analysis of vital signs dynamics: a
case for sepsis onset prediction. Journal of health-
tion, pain, degree of consciousness, and urine output.
care engineering. vol. 2019. 1–12. [Link]
The article highlighted the first five vital indicators
org/10.1155/2019/5930379
as most important. We then examined how hospitals Baidillah, Marlin Ramadhan, Pratondo Busono, and Riyanto
assess patients’ health risks using tracking and trigger- Riyanto. (2023). Mechanical ventilation intervention
ing systems. Most current approaches (typically EWS based on machine learning from vital signs monitor-
or versions thereof) are heuristics with hard-and-fast ing: A scoping review. Measurement Science and Tech-
thresholds, and just a fraction apply AI. The move nology. 34, 062001. Doi: 10.1088/1361-6501/acc11e
from EHRs to PHRs highlights the need of semantic Shengpu, T., Chappell, G. T., Mazzoli, A., Tewari, M., Choi,
interoperability in integrating and exchanging health- S. W., and Wiens, J. (2020). Predicting acute graft-ver-
care data. Today, vital signs may be collected using sus-host disease using machine learning and longitudi-
wearable devices with Bluetooth, NFC, RFID, or nal vital sign data from electronic health records. JCO
Clin. Cancer Inform., 4, 128–135.
UWB connections and gateways to connect hospital
Khan, M. I., Jan, M. A., Muhammad, Y., Do, D.-T., Ur
ward medical equipment. Next, ML processed cru-
Rehman, A., Mavromoustakis, C. X., and Pallis, E.
cial indicators. The IoHT notion introduces various (2021). Tracking vital signs of a patient using channel
issues, allowing for additional research and develop- state information and machine learning for a smart
ment. Prevention and individualization replace symp- healthcare system. Neural Comput. Appl., 1–15.
tom- and disease-focused therapy in the IoHT. This Liu, N. T., Holcomb, J. B., Wade, C. E., Darrah, M. I., and
vital sign monitoring system may help doctors pre- Salinas, J. (2014). Utility of vital signs, heart rate vari-
dict future treatments and interventions. Thus, IoHT ability and complexity, and machine learning for iden-
will improve ward-based patient care. This requires a tifying the need for lifesaving interventions in trauma
patient-centered approach. patients. Shock, 42(2), 108–114.
Li-wei, H. L., Mark, R. G., and Nemati, S. (2016). A model-
based machine learning approach to probing autonom-
References ic regulation from nonstationary vital-sign time series.
Gultepe, Eren, Jeffrey P. Green, and Hien Nguyen. (2013). IEEE J. Biomed. Health Informat., 22(1), 56–66.
From vital signs to clinical outcomes for. 1–11. doi: Li-wei, H. L., Nemati, S., Moody, G. B., Heldt, T., and
10.1136/amiajnl-2013-001815 Mark, R. G. (2014). Uncovering clinical significance
Amit, S., Badotra, S., Bharany, S., Almogren, A., Tag-ElDin, of vital sign dynamics in critical care. Comput. Car-
E. M., and Rehman, A. U. (2022). HealthGuard: An in- diol., 1141–1144.
telligent healthcare system security framework based Srikrishna, I., Zhao, L., Mohan, M. P., Jimeno, J., Siyal,
on machine learning. Sustainability, 14(19), 11934. M. Y., Alphones, A., and Karim, M. F. (2022). mm-
Amit, S., Badotra, S., Rani, S., and Gyaang, R. (2023). Wave radar-based vital signs monitoring and arrhyth-
Evaluation of autism spectrum disorder based on the mia detection using machine learning. Sensors, 22(9),
healthcare by using artificial intelligence strategies. J. 3106.
Sensors, 2023, 1–12. Rana, Soumya Prakash, Maitreyee Dey, Robert Brown,
Amit, S. and Panda, S. N. (2021). Real-time data communi- Hafeez Ur Siddiqui, and Sandra Dudley. (2018). Re-
cation with IoT sensors and things speak cloud. Wire. mote vital sign recognition through machine learning
Sen. Netw. Internet Things, 157–173. augmented UWB. 12th European Conference on An-
Amit, S., Badotra, S., Rani, S., and Gajare, C. M. (2022). tennas and Propagation (EuCAP 2018), London, UK,
WSN-and IoT-based smart surveillance systems for 1–5, doi: 10.1049/cp.2018.0978.
16 Artificial intelligence-based learning techniques for
accurate prediction and classification of colorectal cancer
Yogesh Kumar1,a, Shapali Bansal2, Ankush Jariyal3 and Apeksha Koul4
1
Department of CSE, School of Technology, Pandit Deendayal Energy University, Gandhi Nagar, Gujarat, India
2,3
Department of Computer Applications, USMS, Rayat Bahra University, Mohali, India
4
Department of Computer Science and Engineering, Punjbai University, Patiala, Punjab, India
Abstract
Colorectal cancer (CRC) is a prominent source of illness and death worldwide. Detection and precise diagnosis of CRC at an
early stage can significantly enhance patient outcomes. Artificial intelligence (AI) has yielded promising results in the detec-
tion and classifications of CRC. The application of machine learning (ML) algorithms, deep learning (DL), and computer-
assisted diagnosis systems are only a few of the most current advances in the use of AI techniques for CRC detection and
diagnosis that we discuss in this study. In the article, we also compared and evaluated the CRC detection work of various
researchers using various performance parameters such as accuracy and loss. We also examine the types and epidemiology of
CRC, which aids in the diagnosis of the numerous CRC cancer types. AI has the possible to substantially enhance the detec-
tion and diagnosis of cancer, leading to improved patient health and lower healthcare costs.
Key words: Colorectal cancer, artificial intelligence, epidemiology, deep learning, machine learning, computer-assisted diagnosis
[Link]@[Link]
a
Applied Data Science and Smart Systems 115
The paper is ordered in the following method: Table 16.1 Cases and deaths in the US 2020 due to CRC
Section 2 presents the types and epidemiology and
Age Cases Deaths
various types of CRC. Section 3 presents the con-
(years)
ventional and AI-based diagnosis method. Section 4 CRC Colon Rectum Colorectum
describes the current state-of-the-art techniques for
detecting CRC and highlights any gaps or limitations 0–49 17,930 11,540 6390 3640
in the existing methods. Whereas Section 5 defines the 50–64 50,010 32,290 17720 13,380
methodology and steps to follow the CRC detection 65+ 80.010 60,780 19,230 36,180
using deep learning (DL)-based approaches. Section
All ages 14,7950 10,4610 43,340 53,200
6 concludes the study and presents the significance of
learning models for CRC detection.
II. Types and epidemiology of CRC malignancies are combined and shown in the Table
(American Cancer Society 2020).
There are various types of CRC which are shown in
Table 16.1, along with its brief description and symp-
toms. The CRC can be broadly classified into two III. Diagnosis of CRC
main types: Conventional techniques: People who do physical
activities have been linked to a higher incidence of
Adenocarcinoma: It represents 96% of cases, this is rectal cancer but not colon cancer. According to the
the most prevalent kind of CRC. The cells that lining research, those who are physically active are at the
the inside of the colon and rectum are where adeno- risk of 25% of having distal and proximal colon
carcinoma develops. cancers compared to those who are not. Consuming
Carcinoid tumors: An uncommon form of colon can- aspirin and other non-steroid anti-inflammatory
cer that develop in the intestine’s hormone-producing medicines have also been shown to decrease the risk
cells. Less than 1% of CRCs are caused by them. There of CRC (Howard et al., 2008). Furthermore, other
are also several subtypes of adenocarcinoma of the medications, such as oral bisphosphonates, are used
colon and rectum, which are classified based on their for treating and preventing osteoporosis, which may
microscopic appearance and genetic characteristics. lessen the risk of CRC.
AI techniques: The increasing workload of the pathol-
CRC was rarely identified at least 10 years ago.
ogist in terms of more time and labor consumption
Having 9,00,000 deaths annually is considered the
has tried to incorporate the introduction of computa-
fourth most fatal malignancy globally. It is the most
tional-based pathology for CRC diagnosis. We know
prevalent cancer in men, accounting for 10% of all
that AI has changed the pathology sector. It has been
cases globally, followed by lung cancer (17.2%) and
used to inspect Whole-slide imaging (WSI) data which
prostate cancer (20.3%), and it is the second most
may provide a computer-aided diagnosis of tumors
common cancer in women, accounting for 9.4%
using medical image analysis and various learning
of all cases worldwide, trailing only breast cancer
models such as machine learning (ML) and DL (Cui
(30.9%) (Kanna et al., 2023). In the United States,
et al., 2021).
CRC is the third leading cause of cancer-related mor-
tality among men and women and the second leading Present investigation has demonstrated that AI
cause of cancer deaths among men and women com- plays a vital role to diagnose and treat CRC patients.
bined. It is estimated that 52,580 persons will perish It is a responsible for improving early screening effi-
by 2022 (Sisodia et al., 2023). For several decades, ciency and dramatically improving CRC patients’
the death rate from CRC (per 100,000 persons per 5-year survival rate after treatment. Since 2010, there
year) has decreased in both men and women. There has been a substantial increase in the study and appli-
are several possible explanations for this. One rea- cation of AI in medically assisted gastrointestinal dis-
son is that colorectal polyps are now being discov- ease diagnosis and therapy. AI can help doctors with
ered and removed more frequently through screening the qualitative diagnosis and stage of colon cancer,
before they can develop into malignancies, or cancers which is now reliant on colonoscopy and pathologi-
are being discovered sooner when they are simpler to cal biopsies (Wang et al., 2020).
cure. Furthermore, CRC treatments have improved The researchers, such as Takemura et al. (2012),
during the last few decades (Wolf et al., 2018). utilized narrow-band imaging (NBI) along with a sup-
Table 16.1 projects the number of cases and deaths port vector machine algorithm, a supervised machine
in the United States for 2020. Due to the misclassifica- learning algorithm to calculate extreme points at
tion of rectal cancer deaths as colon, deaths for both the margin. These extreme points were employed to
116 Artificial intelligence-based learning techniques for accurate prediction and classification
identify exceptional parameters on the boundary, histopathology images to review existing research
enabling the differentiation between neoplasia polyps on AI in CRC. According to the authors, DL algo-
and nonneoplasia polyps. The approach achieved a rithms in histopathology are capable of diagnosing,
detection accuracy of 97.8%. identifying the features of histological images related
This shows that an AI can reliably evaluate colo- to prognosis, predicting clinical-based molecular phe-
noscopy biopsies at a rate that is on par with a prac- notypes, and evaluating the specific components of
ticing pathologist. The progress of AI applications the tumor.
in the medical arena suggests that AI will eventually Similarly, (Mitsala et al., 2021) investigated the
be employed for the diagnostic of CRC despite the usefulness of AI systems in medical therapy and diag-
dearth of systematic research. nosis by yielding numerous outstanding outcomes.
They stated that AI-assisted procedures in routine
screening are a critical step in lowering CRC inci-
IV. Related work
dence rates. In this approach, many researchers have
Significant advances have been made by AI techniques used AI algorithms to identify and diagnose CRC, but
in the health arena to demonstrate clinical applica- they also confront significant challenges, as shown in
tion potential. As a result, (Davri et al., 2022) used Table 16.2. This section covers the work done by the
Zhang et al. 2019 1104 endoscopic CNN Accuracy: 86% Class imbalance
non-polyp images, AUC: 1
826 polyp images
Yamada et al. 2019 ImageNet dataset CNN Sensitivity: 97% The system performed
weak in order to detect
AUC: 0.98% lesions in the different
areas of the medical
image
Misawa et al. 2016 1079 narrow band CNN Specificity: 63.3% Unable to classify
imaging images Accuracy: 76.5% correctly because of the
limited dataset
Geetha et al. 2016 703 images Hand Sensitivity: 95% Model trained with
crafted LBP Specificity: 97% limited dataset
Ito et al. 2018 41 cases of colon CNN using Accuracy: 81.2% High cost, low efficiency
endoscopies yielded machine
190 pictures of colon learning
lesions algorithms
Yu et al. 2016 18 colonoscopy CNN Sensitivity: 71% Limited GPU memory,
videos specific length of video
clips were used
Figueiredo 2019 1680 cases of polyps SVM Sensitivity: 99% The model failed to
et al. and 1360 frames of Specificity: 85% evaluate the dimension of
healthy mucosa Accuracy: 91% colorectal polyp
Billah et al. 2017 14,000 still images CNN Prediction rate: Consumes more
98.6% processing time
Ozawa et al. 2020 16,418 images CNN Sensitivity: 92% Less number of training
Accuracy: 83% images were used
Urban et al. 2018 8,641 hand-labeled CNN Sensitivity: 90% The model failed to
images indicate the histology of
polyps
Tsai et al. 2009 CRC-VAL-HE-7K ResNet101 Accuracy: 98.81% Class imbalance issue
researchers to detect CRC using various ML and DL The research methodology for the proposal is men-
techniques along with the research gaps. tioned as under:
• Later, propose a novel hybrid deep learning model rectal cancer using high-resolution MRI. PLoS One,
for the early prediction of different types of CRC. 17(6), e0269931.
• In the end, accuracy, loss, recall, precision, F- Ho, C., Zitong, Z., Xiu, F. C., Jan, S., Sahil, A. S., Rajasa,
score, performance testing, etc., can be used to J., Kaveh, T. et al. (2022). A promising deep learning-
assistive algorithm for histopathological screening of
validate the proposed model’s implemented re-
colorectal cancer. Scientif. Reports, 12(1), 2222.
sults during both the training and testing phase.
Howard, R. A., Michal Freedman, D., Yikyung, P., Albert,
H., Arthur, S., and Michael, F. L. (2008). Physical ac-
VI. Conclusion tivity, sedentary behavior, and the risk of colon and
rectal cancer in the NIH-AARP diet and health study.
The study assists readers (physicians, analysts, doc- Can. causes Con., 19, 939–953.
tors, and so on) in identifying previously utilized Ito, N., Hiroshi, K., Hirotaka, N., Masaya, U., Hideaki,
strategies or algorithms used by the researchers to M., and Hisahiro, M. (2018). Endoscopic diagnostic
detect CRC. In this research, we first focus on the lim- support system for cT1b colorectal cancer using deep
itations of the researchers’ work, such as optimizing learning. Oncol., 96(1), 44–50.
the model and loss function and examine its ability Kanna, G. P., Jagadeesh Kumar, S. J. K., Parthasarathi, P.,
on the significant histopathological dataset. Later, an and Yogesh, K. (2023). A review on prediction and
AI-based model can be designed to assist end users prognosis of the prostate cancer and gleason grading
of prostatic carcinoma using deep transfer learning
in detecting anomalies such as polyps to enhance the
based approaches. Arch. Computat. Methods Engg.,
diagnosis of CRC in a short period. The suggested
1–20.
models can be verified further for real-time images to Koul, A., Rajesh, K. B., and Yogesh, K. (2023). Artificial in-
test its efficacy. telligence techniques to predict the airway disorders
illness: a systematic review. Arch. Computat. Methods
References Engg., 30(2), 831–864.
Koul, A., Yogesh, K., and Anish, G. (2022). A study on
American Cancer Society. (2020). Colorectal cancer facts & bladder cancer detection using AI-based learning tech-
figures 2020–2022. Published Online, 48. niques. 2022 2nd Int. Conf. Technol. Adv. Computat.
Billah, Mustain, Sajjad Waheed, and Mohammad Motiur Sci. (ICTACS), 600–604.
Rahman. (2017). An automatic gastrointestinal pol- Kumar, Y., Inderpreet, K., and Shakti, M. (2023). Food-
yp detection system in video endoscopy using fusion borne disease symptoms, diagnostics, and predictions
of color wavelet and convolutional neural network using artificial intelligence-based learning approaches:
features. International journal of biomedical imag- A systematic review. Arch. Computat. Methods Engg.,
ing. vol. 2017. 1–9. [Link] 1–26.
9545920 Kumar, Y., Surbhi, G., Ruchi, S., and Yu-Chen, H. (2021). A
Chaplot, N., Dhiraj, P., Yogesh, K., and Pushpendra Singh, systematic review of artificial intelligence techniques
S. (2023). A comprehensive analysis of artificial intel- in cancer prediction and diagnosis. Arch. Computat.
ligence techniques for the prediction and prognosis of Methods Engg., 1–28.
genetic disorders using various gene disorders. Arch. Misawa, M., Shin-ei, K., Yuichi, M., Hiroki, N., Shinichi,
Computat. Method Engg., 30(5), 3301–3323. K., Yasuharu, M., Toyoki, K. et al. (2016). Character-
Cui, M. and David, Y. Z. (2021). Artificial intelligence and ization of colorectal lesions using a computer-aided
computational pathology. Lab. Invest., 101(4), 412– diagnostic system for narrow-band imaging endocy-
422. toscopy. Gastroenterol., 150(7), 1531–1532.
Davri, A., Effrosyni, B., Theofilos, K., Georgios, N., Niko- Mitsala, A., Christos, T., Michail, P., Constantinos, S.,
laos, G., Alexandros, T. T., and Anna, B. (2022). Deep and Alexandra, K. T. (2021). Artificial intelligence in
learning on histopathological images for colorectal colorectal cancer screening, diagnosis and treatment.
cancer diagnosis: A systematic review. Diagnostics, A new era. Curr. Oncol., 28(3), 1581–1607.
12(4), 837. Ozawa, T., Soichiro, I., Mitsuhiro, F., Youichi, K., Satoki,
Fearon, E. R. (2011). Molecular genetics of colorectal can- S., and Tomohiro, T. (2020). Automated endoscopic
cer. Ann. Rev. Pathol. Mec. Dis., 6, 479–507. detection and classification of colorectal polyps using
Figueiredo, P. N., Isabel, N. F., Luís, P., Sunil, K., Yen-his, convolutional neural networks. Ther. Adv. Gastroen-
R. T., and Alexander, V. M. (2019). Polyp detection terol., 13, 1756284820910659.
with computer-aided diagnosis in white light colonos- Sisodia, P. S., Gaurav, K. A., Yogesh, K., and Neelam, C.
copy: comparison of three different methods. Endo. (2023). A review of deep transfer learning approaches
Int. Open, 7(02), E209–E215. for class-wise prediction of Alzheimer’s disease using
Geetha, K. and Rajan, C. (2016). Automatic colorectal pol- MRI images. Arch. Computat. Methods Engg., 30(4),
yp detection in colonoscopy video frames. Asian Pac. 2409–2429.
J. Can. Preven. APJCP, 17(11), 4869. Takemura, Y., Shigeto, Y., Shinji, T., Rie, K., Keiichi, O.,
Hamabe, A., Masayuki, I., Rena, K., Saeko, S., Koichi, O., Shiro, O., Toru, T. et al. (2012). Computer-aided sys-
Kenji, O., Emi, A. et al. (2022). Artificial intelligence– tem for predicting the histology of colorectal tumors
based technology for semi-automated segmentation of by using narrow-band imaging magnifying colonos-
Applied Data Science and Smart Systems 119
copy (with video). Gastrointes. Endoscop., 75(1), Colorectal cancer screening for average-risk adults:
179–185. 2018 guideline update from the American Cancer So-
Tsai, H.-L., Koung-Shing, C., Yu-Ho, H., Yu-Chung, S., ciety. CA Can. J. Clin., 68(4), 250–281.
Jeng-Yih, W., Chao-Hung, K., Chao-Wen, C., and Jaw- Yamada, M., Yutaka, S., Hitoshi, I., Masahiro, S., Shigemi,
Yuan, W. (2009). Predictive factors of early relapse in Y., Hiroko, K., Hiroyuki, T. et al. (2019). Development
UICC stage I–III colorectal cancer patients after cura- of a real-time endoscopic image diagnosis support sys-
tive resection. J. Surg. Oncol., 100(8), 736–743. tem using deep learning technology in colonoscopy.
Urban, G., Priyam, T., Talal, A., Mohit, M., Farid, J., Wil- Scientif. Reports, 9(1), 14465.
liam, K., and Pierre, B. (2018). Deep learning local- Yu, L., Hao, C., Qi, D., Jing, Q., and Pheng, A. H. (2016).
izes and identifies polyps in real time with 96% accu- Integrating online and offline three-dimensional deep
racy in screening colonoscopy. Gastroenterol. 155(4), learning for automated polyp detection in colonosco-
1069–1078. py videos. IEEE J. Biomed. Health Informat., 21(1),
Wang, Y., Xiaoyun, H., Hui, N., Jianhua, Z., Pengfei, C., 65–75.
and Chunlin, O. (2020). Application of artificial in- Zhang, X., Yang, Y., Yalan, W., and Qi, F. (2019). Detection
telligence to the diagnosis and therapy of colorectal of the BRAF V600E mutation in colorectal cancer by
cancer. Am. J. Can. Res., 10(11), 3575. NIR spectroscopy in conjunction with counter propa-
Wolf, A., Elizabeth, T. H. F., Timothy, R. C., Christopher, R. gation artificial neural network. Molecules, 24(12),
F., Carmen, E. G., Samuel, J. L., Ruth, E. et al. (2018). 2238.
17 SLODS: Real-time smart lane detection and object
detection system
Tanuja Satish Dhope1,a, Pranav Chippalkatti2, Sulakshana Patil3,
Vijaya Gopalrao Rajeshwarkar3 and Jyoti Ramesh Gangane4
1
Department of Electronics and Communication, Bharati Vidyapeeth (Deemed to be University) College of Engineer-
ing, Pune, Maharashtra, India
2
Department of Computer Science and Engineering, School of Computing, MIT Art, Design and Technology Univer-
sity, Pune, Maharashtra, India
3
Department of Electronics and Telecommunication, Sinhgad Institute of Technology, Lonavala, Pune, Maharashtra,
India
4
Department of Electronics and Telecommunication, Vishwaniketan’s Institute of Management Entrepreneurship and
Engineering Technology, India
Abstract
With the advances in technologies, autonomous cars/self-driving cars are now-a-days gaining more demand due to the in-
crement in mortality rate by road accidents caused due to human errors. Detecting obstacles on a road is one of the biggest
challenges in autonomous vehicle/self-driving navigation system. In this paper, we have proposed the real-time smart lane
detection and object detection system (SLODS) which captures the real-time road traffic using two cameras, one in the front
and the other one at the back of the car. The front one detects the lane while the other one detects if any other vehicle is ap-
proaching while changing the lanes, ensuring safe lane change. Region of interest (ROI) determines object and lane detection.
The performance of the edge detection algorithms like Roberts, Sobel, Prewitt’s, and Canny edge detectors, are evaluated
based on precision, recall, F1 score, and peak signal to noise ratio (PSNR) values. For PSNR, Canny is outperforming oth-
ers by the difference of -39dB with Sobel, -14 dB with Prewitt, and -48 dB. Further the proposed system also calculates the
speed of the approaching vehicle.
Keywords: Lane detection, edge detection, object detection, machine learning, Hough transform
tanuja_dhope@[Link]
a
Applied Data Science and Smart Systems 121
S. No. Vertex X Y
1 Rho 1
2 Beta π / 180
3 Minimum votes 15
threshold
4 Minimum line 7
Figure 17.3 Lane detection flowchart length
5 Maximum line 3
gap
β = angle between of x-axis along with vector. 6 Line thickness 1
A flowchart showing the sequence of events that
take place during lane detection is shown in Figure
17.3.
Steps for lane detection according to the Figure 17.1: 6. Applying Hough transform: Hough transform is
a technique to extract features from the image
1. Reading video and dividing into frames: The in- to analyze it. It is extensively used for image an-
put footage is recorded by a camera mounted alyzation based on shapes like rectangles, circles,
on the vehicle. The video input is then divided etc. It assumes all the white pixels of the image to
into frames (images) which are used to determine be the points and converts them into ρ-β plane.
lanes and boundaries on a laned road or high- ρ line connects polar coordinates to the origin
way. where the x-axis intersects the y-axis (Peerawat
2. Converting image into gray scale form: This is Mongkonyong et al., 2018) (Table 17. 2).
done to avoid recording unnecessary pixels. As a
result, far less information is collected and evalu- (2)
ated as compared to a colorful image.
3. Noise reduction using filter: The image produced Figures 17.4–17.8 gives a detailed idea of the step
after gray scale conversion is of poor quality. To 1–5 performed by the algorithms.
improve the accuracy of the recognized items,
noise filtration from the gray scaled image is B. Edge detection techniques
done before applying ML algorithm for object An edge is defined as an area of significant change
detection. in image intensity/contrast. Locating the areas with
4. Detecting edges: One of the topic’s cornerstones great intensity contrasts is called edge detection. Now
is edge detection. To detect a picture, the input it’s also possible that a certain pixel can accommo-
gray scaled image is exposed to the various edge date any variation and we can mistake it for an edge.
detectors (discussed in next section) after being Different situations can lead to this, for example, in
filtered (Zakir Hussain et al., 2015; Zhi Zhang et low light conditions or there can be noise that can
al., 2016; VanQuang Nguyena et al., 2018). This show all the characteristics of edge color segmentation.
gives us the power to modify the intensity of the
frames. B.1 Robert’s operator
5. Choosing region of interest (ROI): The area Robert’s operator (Zakir Hussain et al., 2015; Zhi
of the image in which the lane is detected and Zhang et al., 2016; VanQuang Nguyena et al., 2018)
placed in an area referred to as ROI. In our sys- is a type of an operator that works using cross prod-
tem ROI is taken as follows (Table 17.1). ucts to determine the grade of the detected image
Applied Data Science and Smart Systems 123
(3)
(4)
Figure 17.5 Frame gray scaling
(5)
Figure 17.6 Denoising Now we can also calculate the gradient direction
given by Equation 6
(6)
(7)
(8)
(9)
(10)
1 Low threshold 50
2 High threshold 150
Figure 17.12 Frame gray scaling
(11)
PSNR
Image Pixel Precision Recall F1 score (dB)
F1 PSNR
Table 17.4 Parameters for Robert operator Image Pixel Precision Recall score (dB)
Below is the comparison table for all detector’s that Canny outperforms for the various images in
algorithms for five images extracted from real time terms of other parameters also.
video. For lane detection, on an empty road, with a
The different edge detecting operators were tested straight lane, we apply Hough transform on the
based on various parameters like Precision, Recall, F1 detected edges using Canny edge detector. The lanes
score and PSNR values. Both Sobel and Prewitt edge on the road are detected and then highlighted using
detector were able to detect edges successfully, but the orange color markings (see Figure 17.16). Thus, even-
number of edges detected were far lower than Canny tually making it easier for the driver to navigate on
edge detection method (see Tables 17.5–17.7). For the road.
example, PSNR provided by Robert, Sobel, Prewitt, As soon as the incoming vehicle enters the specified
Canny is 66 dB, 40 dB, 70 dB and 82 dB for img ROI, the object detection algorithm starts detecting
1, respectively. Apart from low processing time, the the vehicle and it is finally bounded in a bounding
canny operator also displays a higher precision rate box (see Figure 17.17). Thus, the driver is alerted for
and better PSNR as compared to other operators. safe lane change.
Also, F1 score provided by Canny is 0.91 compared To detect the speed of oncoming vehicle ROI is
to others for img 1. Further Tables 17.5–17.7 indicate used. ROI has been set up as a distance up to 80 m
Applied Data Science and Smart Systems 127
References
Annual report on Road Accidents in India. (2020). Re-
trieved from https:// [Link]/sites /default/files/
RA_2020.pdf.
Madhura, B., Tanuja, D., Akshay, V., and Dina, S. (2022).
Figure 17.17 Real time object detection Performance analysis of object classification system
for traffic objects using various SVM Kkernels. Adv.
Data Comput. Comm. Sec., 423–432, doi:[Link]
org/10.1007/978-981-16-8403-6_39.
Table 17.8 Speed analysis
Sandipann, N., Pradnya, N. B., and Dhiraj, M. D. (2018). A
Object Distance (m) Speed (m/s) review of recent advances in lane detection and depar-
ture warning system. Pattern Recogn., 73, 216–234.
Object 1 70 35 doi:[Link] 10.1016/ [Link].2017.08.014.
Anuj, M., Constantine, P., and Tomaso, P. (2001). Exam-
Object 2 75 27.5
ple-based object detection in images by components.
Object 3 60 30 IEEE Trans. Pattern Anal. Mac. Intel., 23(4), 349–
Object 4 65 37.5 361. doi:https: //[Link]/ 10.1109/ 34.917571.
Bertozzi , M. and Broggi, A. (1998). GOLD: A parallel real-
Object 5 72 36
time stereo vision system for generic obstacle and lane
detection. IEEE Trans. Image Proc., 7(1), 62–81. doi:
[Link] 83.650851.
Mukesh, T. and Rakesh, S. (2017). A review of detec-
which can be varied according to user requirement. tion and tracking of object from image and video
The speed at which the test vehicle is moving is 25 sequences. Int. J. Computat. Intel. Res., 13(5),
m/s. A time interval of 2 s has been chosen for a vehi- 745–765. [Link]
cle to cover its distance (see Table 17.8). ijcirv13n5_07 .pdf.
William, N., Jack, L., Simon, G., and Jaco, V. (2005). A re-
view of recent results in multiple target tracking. Proc.
VI. Conclusion 4th Int. Sym. Image Sig. Proc. Anal., 3807–3812.
Four distinct edge detection algorithms are men- doi:[Link] 10.1109/ ISPA.2005. 195381.
tioned in the study for our proposed SLODS. PSNR Sunil, K., Vishwakarma, A., and Divakar, S. Y. (2015).
Values, Precision, Recall, and F1 Score were the Analysis of lane detection techniques using OpenCV.
2015 Ann. IEEE India Conf., 1–4. doi:[Link]
metrics utilized to assess these approaches. These
org// 10.1109/ INDICON. 2015.7443166.
parameters were utilized in this study to assess the
Xu, Y. and Zhang, L. (2015). Research on lane detec-
performance of the edge detection approaches devel- tion technology based on Opencv. 3rd Int. Conf.
oped by Canny, Sobel, Prewitt, and Robert. After Mech. Engg. Intel. Sys., 994–997. doi: [Link]
careful examination, we determined that the edge org//10.2991/icmeis-15.2015.187.
detection approach successfully detected the great- Zhong-Xun, W. and Wenqi, W. (2018). The research on edge
est number of edges for both vertical and horizontal detection algorithm of lane. EURASIP J. Image Video
edges. If we visualize images in Figure 17.9, we can Proc., 98, 1–9. doi: [Link]
clearly see that Robert, Prewitt and Sobel give a low- 018-0326-2.
quality image as output when compared to Canny. Xining, Y., Dezhi, G., Jianmin, D., and Lei, Y. (2011). Re-
The Canny method on the other hand can detect search on lane detection based on machine vision. Proc.
2011 Int. Conf. Informat. Cybernet. Comp. Engg.,
both weak and strong edges. The paper also men-
110, 539–547. doi: [Link]
tions a method to detect and track objects in an effi-
642-25185-6_69.
cient manner. It suggests selecting and then masking VanQuang, N., Heungsuk, K., SeoChang, J., and Kwang-
the ROI after gray scale conversion of the image. The suck, B. (2018). A study on real-time detection method
algorithm used in this paper was able to detect 96% of lane and vehicle for lane change assistant system
of all the vehicles in the image. The SLODS provides using vision system on highway. Engg. Sci. Technol.
accurate speed estimation of the oncoming vehicle as Int. J., 21(5), 822–833. doi: [Link] 10.1016/
well. [Link].2018.06.006.
128 SLODS: Real-time smart lane detection and object detection system
Singh, J., Singh, S., Singh, S., and Singh, H. (2019). Evaluat- transform. IOP Conf. Ser. Mat. Sci. Engg., 297(1),
ing the performance of map matching algorithms for 1–11. doi: [Link] article/ 10.1088
navigation systems: an empirical study. Spat. Inform. /1757-899X/297/1/012050.
Res., 27, 63–74. Assidiq, A. A., Khalifa, O. O., Islam, M. R., and Khan, S.
Hussain, Zakir, and Diwakar Agarwal. (2015). A com- (2008). Real time lane detection for autonomous ve-
parative analysis of edge detection techniques used in hicles. Int. Conf. Comp. Comm. Engg., 82–88. doi:
flame image processing. International Journal of Ad- [Link]
vance Research In Science And Engineering IJARSE, Ziqiang, S. (2020). Vision based lane detection for self-
4, 3703–3711. driving car. 2020 IEEE Int Conf Adv. Elec. Engg.
Zhi, Z., Zhihai, H., Guitao, C., and Wenming, C. (2016). Comp. Appl., 635–638. doi: [Link]
Animal detection from highly cluttered natural AEECA49918.2020.9213624.
scenes using spatiotemporal object region pro- John, C. (1986). A computational approach to edge de-
posal sand patch verification. IEEE Trans. Mul- tection. IEEE Trans. Pat. Anal. Mac. Intel., 8(6),
timed., 18(10), 1–14. doi: [Link] 679–698. doi: https:// [Link] /10.1109/TPAMI.1986.
TMM.2016.2594138. 4767851.
Peerawat, M., Chaiwat, N., Supakorn, S., and Masaki, Y.
(2018). Lane detection using randomized Hough
18 Computational task off-loading using deep Q-learning in
mobile edge computing
Tanuja Satish Dhopea, Tanmay Dikshit, Unnati Gupta and Kumar Kartik
Department of Electronics and Communication, Bharati Vidyapeeth (Deemed to be University) College of Engineering,
Pune, India
Abstract
Because of the growing proliferation of networked Inter of Things (IoT) devices and the demanding requirements of IoT
applications, existing cloud computing (CC) architectures have encountered significant challenges. A novel mobile edge com-
puting (MEC) can bring cloud computing capabilities to the edge network and support computationally expensive applica-
tions. By shifting local workloads to edge servers, it enhances the functionality of mobile devices and the user experience.
Computation off-loading (CO) is a crucial mobile edge computing technology to enhance the performance and minimize
the delay. In this paper, the deep Q-learning method has been utilized to make off-loading decisions whenever numerous
workloads are running concurrently on one user equipment (UE) or on a cellular network, for better resource management in
MEC. The suggested technique determines which tasks should be assigned to the edge server by examining the CPU utiliza-
tion needs for each task. This reduces the amount of power and execution time needed.
Keywords: Computation off-loading, edge server, mobile edge server, deep Q-learning
a
tanuja_dhope@[Link]
130 Computational task off-loading using deep Q-learning in mobile edge computing
connections to become severely congested if many best application approach while adhering to work-
application stations forcefully dump their processing flow applications deadline constraints. Numerous
resources to the edge node, which would dramatically experimental evaluations have been carried out to
slow down MEC (Gagandeep Kaur et al., 2021). A demonstrate the usefulness and efficiency of suggested
unified management system for CO and the accom- strategy.
panying wireless resource distribution in order to In MEC wireless networks, an Software Defined
benefit from compute off-loading is required (Khadija Networking (SDN) -based solution for off-loading
Akherfi et al., 2016). In section 2, this paper describes compute. Based on reinforcement learning, a solu-
the job off-loading research in MEC. The task off- tion to the energy conservation problem that consid-
loading system model is described in section 3 as local ers both incentives and penalties have been assessed
computing, edge computing, and the deep Q-learning (Nahida Kiran et al., 2020).
method. Section 4 elaborates on the results and charts Distributed off-loading method with deep rein-
for various task off-loading techniques. Finally con- forcement learning that allows mobile devices to
clusion is presented in section 5. make their off-loading decisions in a decentralized
way has been proposed. Simulation findings dem-
II. Related work onstrated that suggested technique may decrease the
ratio of dropped jobs and average latency when com-
In (Khadija Akherfi et al., 2016) many edge com- pared to numerous benchmark methods (Ming Tang
puting paradigms and their various applications, as et al., 2020). A multi-layer CO optimization frame-
well as the difficulties that academics and industry work appropriate for multi-user, multi-channel, and
professionals encounter in this fast-paced area has multi-server situations in MEC has been suggested.
been examined. Author suggested options, including Energy consumption and latency parameters are used
establishing a middleware-based design employing an for CO decision from the perspective of edge users.
optimizing off-loading mechanism, which might help Multi-objective decision-making technique has been
to improve the current frameworks and provide the proposed to decrease energy consumption and delay
mobile cloud computing (MCC) users more effective of the edge client (Nanliang Shan et al., 2020).
and adaptable solutions by conserving energy, speed-
ing up reaction times, and lowering execution costs. III. Methodology
Ke Zhang et al. (2016) has given an energy efficient
computation off-loading (EECO) method, which We took into account energy-sensitive UEs in this
combinedly optimizes the decisions of CO and alloca- paper, such as IoT devices and sensor nodes, which
tion of radio resources thereby minimizing the cost have low power requirements but are not delay-sen-
of system energy within the delay constraints in 5G sitive. We take into account N energy-sensitive UEs
heterogeneous networks. that are running concurrently on a server, and the
An energy-efficient caching (EEC) techniques for a server must choose which task from the task queue
backhaul capacity-limited cellular network to reduce needs to be done first in order to reduce the power
power consumption while meeting a cost limitation and execution time for each work. When a user device
for computation latency has been proposed (Zhaohui lacks the energy resources to complete the computa-
Luo et al., 2019; Gera et al., 2021). The numerical tion-intensive task locally, an edge server can step in.
findings demonstrate that 20% increase in delay effi- To make decisions on off-loading, we employ deep
ciency. The proposed method may be very close to the Q-learning algorithm. Based on the state and reward
ideal answer and far superior to the most likely out- of the Q function at state t, the Q-learning algorithm
come, i.e., the approximation bound. acts.
With the use of two time-optimized sequential
decision-making models and the optimal stopping A. Local computing
theory, author (Ibrahim Alghamdi et al., 2019) Let’s assume that E represents the energy needed for
address the issue of where to off-load from and when each User Equipment (UE) to operate locally n= num-
to do so. Real-world data sets are used to offer a ber of UE, pn =the power coefficient of energy used
performance evaluation, which is then contrasted for local computing per CPU cycle, cn= the CPU cycles
with baseline deterministic and stochastic models. desired for each bit in numbers, βn = the percentage of
The outcomes demonstrate that, in cases involving tasks computed locally, and Sn = the size of the com-
a single user and rival users, our technique optimizes putation task are all represented by the numbers n.
such decisions. Therefore, the amount of energy needed for UE to
Kai Peng et al. (2019) examine the multi-objective operate locally can be determined by discretion.
computation off-loading approach for workflow
applications (MCOWA) in MEC which discovers the (1)
Applied Data Science and Smart Systems 131
Parameters Value
Abstract
Drowsiness or tiredness is a leading cause of accidents on the road, posing a serious threat to safety. Many accidents can
be prevented if drowsy drivers can be alerted within time. Several drowsiness detection methods are available to observe
drivers’ alertness during their journey and alert them if they are distracted. These methods gauges drowsiness by looking for
signs like yawning, closed eyes, or unusual head movements. They also consider the driver’s physical condition and vehicle
behavior. This paper offers a comprehensive analysis of the existing drowsiness detection methods and a detailed review of
the common classification techniques. It first categorizes the current methods into those based on subjective, behavioral, ve-
hicular, and physiological parameters based. Finally, it examines the strengths and weaknesses of these various methods and
compares them. In conclusion, the paper summarizes the research findings from this comprehensive survey to guide other
researchers toward potential future work in this field.
[Link]@chitkara..[Link]
a
Applied Data Science and Smart Systems 135
2.2.2 Mouth and Yawning analysis compared with those obtained from testing images to
Yawning, often a result of tiredness or boredom, determine the appropriate classification. The system
can signal a potential risk for drivers, suggesting computes the duration of closed eyes and identifies
they might doze off while driving. Techniques exist drowsiness if this duration surpasses a predefined
to gauge the extent of mouth widening, serving as a time. Furthermore, it evaluates various head move-
means to detect signs of yawning in drivers. ments such as left, right, forward, backward, and
Yan et al. (2016) proposed an effective method for tilting motions in both directions. To conduct this
monitoring driver fatigue using Yawning extraction. analysis, the video footage is divided into frames,
To begin, the method uses the support vector machine with the system examining head images and com-
(SVM) technique to extract the face region from paring their positions to determine head postures.
images, reducing associated costs. The method pro- Subsequently, it merges the duration of closed eyes
ceeds to locate the mouth: facial edges are detected with the assessment of head positions to ascertain
through an edge detection technique, followed by drowsiness. The methodology was tested using six
a vertical projection to determine the right and left videos simulating genuine driving conditions and the
boundaries in the lower face area. Then, a horizon- findings are displayed through a confusion matrix. It
tal projection helps identify the upper and lower achieved a 98% accuracy rate, proving more effective
mouth limits, defining the mouth’s localized region. than alternative detection methods.
For yawning detection, the system employs circular The main problem with vision-based approach
Hough transform (CHT) on mouth region images is lighting. Regular cameras struggle at night.
to spot wide-open mouths. An alert is generated if Furthermore, many methods have been tested using
a notable number of consecutive frames capture a data from drivers imitating drowsiness instead of
widely open mouth. The method’s effectiveness is com- using real videos capturing a driver naturally becom-
pared with various edge detectors like Sobel, Roberts, ing sleepy.
Prewitt, and Canny. The experiment utilizes six videos
simulating real driving conditions, and the results are 2.3 Vehicle-based measures
depicted in a confusion matrix. The proposed method Vehicle-based measures detect driver fatigue using
attains a 98% accuracy rate, surpassing the perfor- vehicular features, including steering wheel angle,
mance of all other edge detection techniques. steering wheel grip force, lane changing patterns, and
vehicle speed variability. These measures necessitate
2.2.3 Head position the installation of sensors on various vehicle compo-
The head’s position is another sign of tiredness and nents, such as the steering wheel, accelerator, or brake
drowsiness in drivers. When feeling drowsy, drivers pedal, among others. The signals produced by these
often tend to tilt, lower, or nod their heads, particu- sensors serve as the basis for evaluating the drowsi-
larly in the later stages of sleepiness. Several factors, ness levels of drivers.
including decreased muscle tone, reduced vigilance,
and brief periods of sleep, can cause these changes in 2.3.1 Lane detection
the head position. Monitoring head position is a nota- This approach checks the vehicle’s position with
bly effective method for identifying driver drowsiness, respect to the middle of the lane. It is also known as
given that it is relatively easy and inexpensive to mon- the standard deviation of lane position (SDLP). Katyal
itor. Moreover, the head position remains unaffected et al. (2014) proposed a driver’s drowsiness detection
by environmental elements like lighting or noise. system using lane and driver’s fatigue level. Hough
Teyeb et al. (2014) proposed a method for drowsy transform is used to detect lanes and canny edge
driver detection using eye closure and head postures. detection is applied over viola-jones to detect eyes and
The system begins by capturing video through a web- driver’s fatigue level. This information is then used to
cam and conducting following operations on each detect improper driving. Ingre et al. (2006) conducted
frame of the video. It employs the Viola-Jones method multiple experiments and concluded that KSS ratings
to identify the ROI, encompassing the face and eyes. are directly proportional to SDLP metrics.
Subsequently, the facial area is sub-divided into sec-
tions, and the Haar classifier is utilized to focus on 2.3.2 Steering wheel analysis
the upper segment, specifically targeting the region Steering wheel analysis (SWA) is a widely used
corresponding to the eyes for analysis. Following this, vehicle-based measure to detect driver drowsiness
identifying the eye state entails the utilization of a (Fairclough et al., 1999; Thiffault, 2003). An angle
Wavelet network, a neural network-based approach, sensor is attached on the steering wheel axis to collect
which is trained using image data. The learning pro- the data. Abnormal steering wheel reversals, steering
cess involves ascertaining coefficients from training correction periods, and a vehicle’s jerky motion indi-
images. These learned coefficients are subsequently cate fatigue and a drowsy driver. Li et al. (2017) uses
Applied Data Science and Smart Systems 137
SWA and proposed an online drowsiness detection apt for detecting drowsiness. Leveraging physiologi-
system to monitor the fatigue level of drivers under cal signals for drowsiness detection holds the poten-
natural conditions by extracting approximate entropy tial to mitigate the issue of false positives, which
features and using a decision classifier for detection. is a common challenge with existing approaches.
Zhenhai et al. (2017) proposed a solution by analyz- Furthermore, it enables timely alerts, thereby averting
ing the time series of the angular velocity of the steer- road accidents.
ing wheel. Fairclough and Graham (1999) proposed a
solution by checking the steering wheel’s reversals and 2.4.1 Electroencephalography (EEG)
small SWMs. They found that drowsy drivers make EEG measures the brain’s electrical activity by plac-
fewer steering wheel reversals than typical drivers. ing some electrodes on the head and forehead. The
Many studies have shown that vehicle-based mea- frequency of signals ranges from 1 to 50 Hz and
sures are not the best way to judge a driver’s drowsiness amplitude from 20 to 200 μV. Some frequency bands
and often lead to inaccurate results. Assessing driver are defined as alpha waves (8–12 Hz, 25–100 μV),
fatigue solely based on vehicle movement has limita- which measure relaxation; beta waves (faster than 13
tions, as the measurement metrics can be susceptible Hz and below 40 μV), which measure alertness; theta
to external influences like the road’s geometric attri- waves (4–7 Hz, 20–120 μV), which measure drowsi-
butes and prevailing weather conditions. Other factors ness; and delta waves (0.5–3.5 Hz, 75–200 μV) helps
can also affect these measures, such as road, traffic, to check if the subject is asleep.
lighting conditions, and driving under the influence of Several studies support the connection between
alcohol or other drugs. Steering wheel grip force on a EEG signals and driver behavior (Campagne et al.,
curvy mountain road differs significantly from that of 2004; Akin et al., 2008; Liu et al., 2010; Lin et al.,
a straight highway. Furthermore, the driver’s grip can 2012; Lin et al., 2013). Changes in the alpha frequency
vary with road conditions. Driver may not grip the band, where the power decreases, and an increase in
steering with that pressure on a busy road with which the theta frequency band are indicative of drowsiness.
he grips the steering on an empty expressway. Akin et al. (2008) observed that combining EEG and
EMG signals is more successful in detecting drowsi-
2.4 Physiological measures ness compared to using either signal alone.
As drivers experience fatigue, they may observe a
subtle swaying of their heads, and there is an elevated 2.4.2 Electrocardiography (ECG)
risk of the vehicle deviating from the center of the The ECG method records the heart’s electrical activity
lane. Previously discussed methods for detecting this by positioning electrodes on the chest, arms, and legs,
behavior, such as behavior-based and vehicle-based, capturing the small electrical signals generated with
possess limitations, primarily because they can only each heartbeat.
detect fatigue after the driver has already entered a Tsuchida et al. (2009) research claims that heart
drowsy state. rate variability (HRV) can be used to detect driver
However, it is worth noting that physiological sig- fatigue and drowsiness. As drivers get tired, their
nals undergo discernible changes early in the onset parasympathetic activity decreases, and their sympa-
of drowsiness. Hence, physiological signals are more thetic activity increases. This causes a notable shift in
138 A comprehensive analysis of driver drowsiness detection techniques
cardiac rhythm from a high-frequency range of 0.15– Table 19.1 List of various work done on driver drowsiness
0.4 Hz to a lower frequency range of 0.04–0.15 Hz. detection.
Several studies have explored driver fatigue and
S. Measure Method Algorithm Accuracy
drowsiness detection using photo plethysmogram No. used (%)
(PPG) and electrocardiogram (ECG) wavelet spec-
trum analysis. Tsuchida et al. (2009), Arun et al. 1 Behavioral Eye-blink Viola Jones 94
(2012), Lee et al. (2014), reporting an average predic- rate
tion accuracy of 96% in their experimental findings. 2 Behavioral Yawning SVM 98
analysis
2.4.3 Electromyography (EMG) 3 Behavioral Head Viola Jones 98
EMG measures the electrical activity of muscles and position with Haar
is commonly obtained from the chin (Hostens, 2005). classifier
When a muscle contracts, it sends electrical signals to 4 Physiological PPG and 96
the brain. ECG
Katsis et al. (2004) observed up to 20% frequency 5 Physiological, EOG Haar 80
decrease and up to 50% amplitude increase after behavioral with eye classifier
movement
monotonous driving tasks and used them to indicate
fatigue and drowsiness. Balasubramanian et al. (2007) 6 Physiological, Heart PERCLOS 96
behavioral rate with
also had similar observations in EMG from shoulder eyelid
and neck muscles during 15 min of simulated driving. closure
ratio
2.4.4 Electrooculography (EOG)
EOG measures the electrical potential difference
between the human eye’s front (cornea) and back (ret-
elucidated, and their advantages and drawbacks are
ina). It’s one of the primary functions is to gauge the
considered. Nonetheless, certain gaps have been pin-
amplitude and the direction of eye movements, which
pointed in the existing literature, such as the need to
is applicable in detecting driver drowsiness (Shuvan et
evaluate current techniques in real time. This is par-
al., 2009). The electric potential difference between the
ticularly crucial in dynamic driving conditions and
retina and cornea generates an electrical field which is
diverse environmental factors, presenting opportuni-
measured using EOG sensors and determines eyes ori-
ties for refinement and enhancement. A comparative
entation. By employing a disposable Ag–Cl electrode
analysis reveals that no single method achieves abso-
on each eye’s outer corner and a third electrode at
lute accuracy, although techniques relying on physi-
the forehead’s center, the system observes horizontal
ological parameters tend to yield more precise results
eye movements (Shuvan et al., 2009). These electrodes
than others. A combination of these methods, includ-
assist in identifying behavioral patterns such as rapid
ing physiological, vehicular, or behavioral measures,
eye movements (REM) and slow eye movements
can address the limitations present in each technique
(SEM), contributing to drowsiness detection in driv-
when used individually (Table 19.1).
ers (Lal et al., 2001; Sharma et al., 2020).
Khushaba et al. (2010) and Kukreja et al. (2022)
discovered that EOG alone could not produce accu- References
rate results compared to EEG alone for detecting Sales Statistics | [Link]. (n.d.). Accessed September
drowsiness. Chieh et al. (2005) monitored eye move- 17, 2023. [Link]
ment using EOG rather than a video-based eye moni- World Health Organization: WHO. (2022). Road Traffic
tor and achieved 80% accuracy. Injuries. 2022. [Link]
sheets/detail/road%09traffic-injuries.
National Highway Traffic Safety Administration. (n.d.). Ac-
III. Result and discussion cessed September 17, 2023. [Link]
The issue of driver drowsiness represents a significant [Link]/#!/.
threat to road safety. Detecting driver drowsiness and McKernon, S. (2009). A literature review on driver fatigue
promptly issuing alerts is essential to avert a substan- among drivers in the general public. Land Transport,
New Zealand. 1–62.
tial number of road accidents. The primary objective
Liu, C. C., Simon, G. H., and Michael, G. L. (2009). Predict-
of this systematic review is to explore the most cur-
ing driver drowsiness using vehicle measures: Recent
rent advancements in drowsiness detection systems. insights and future challenges. J. Safety Res., 40(4),
This review examines drowsiness detection methods 239–245.
based on subjective, behavioral, vehicular, and physi- Brown, T., John, L., Chris, S., Dary, F., and Anthony, M.
ological parameters. These methods are thoroughly (2014). Assessing the feasibility of vehicle-based
Applied Data Science and Smart Systems 139
sensors to detect drowsy driving. No. DOT HS 811 ness prediction system by using a self-organizing neu-
886. ral fuzzy system. IEEE Trans. Circuit. Sys. I Reg. Pa-
Sigari, M.-H., Muhammad-Reza, P., Mohsen, S., and Mah- pers, 59(9), 2044–2055.
mood, F. (2014). A review on driver face monitoring Liu, J., Chong, Z., and Chongxun, Z. (2010). EEG-based
systems for fatigue and distraction detection. Int. J. estimation of mental fatigue by using KPCA–HMM
Adv. Sci. Technol., 64, 73–100. and complexity parameters. Biomed. Sig. Proc. Con.,
Mittal, A., Kanika, K., Sarina, D., and Manvjeet, K. (2016). 5(2), 124–130.
Head movement-based driver drowsiness detection: Campagne, A., Thierry, P., and Alain, M. (2004). Correlation
A review of state-of-art techniques. 2016 IEEE Int. between driving errors and vigilance level: influence of
Conf. Engg. Technol. (ICETECH), 903–908. the driver’s age. Physiol. Behav., 80(4), 515–524.
Sanjaya, K. H., Soomin, L., and Tetsuo, K. (2016). Review Lin, C.-T., Kuan-Chih, H., Chun-Hsiang, C., Li-Wei, K., and
on the application of physiological and biomechanical Tzyy-Ping. J. (2013). Can arousing feedback rectify
measurement methods in driving fatigue detection. J. lapses in driving? Prediction from EEG power spectra.
Mechat. Elec. Power Vehicul. Technol., 7(1), 35–48. J. Neural Engg., 10(5), 056024.
Sahayadhas, A., Kenneth, S., and Murugappan, M. (2013). Tsuchida, A., Md Shoaib, B., and Koji, O. (2009). Estima-
Drowsiness detection during different times of day us- tion of drowsiness level based on eyelid closure and
ing multiple features. Aus. Phy. Engg Sci. Med., 36, heart rate variability. 2009 Ann. Int. Conf. IEEE
243–250. Engg. Med. Biol. Soc., 2543–2546.
Kang, H.-B. (2013). Various approaches for driver and driv- Arun, S., Kenneth, S., and Murugappan, M. (2012). Hypo-
ing behavior monitoring: A review. Proc. IEEE Int. vigilance detection using energy of electrocardiogram
Conf. Comp. Vis. Workshops, 616–623. signals. Journal of scientific & Industrial Research.
Rahman, A., Mehreen, S., and Aliya, K. (2015). Real time 71(12), 794–799.
drowsiness detection using eye blink monitoring. 2015 Lee, B.-G., Jae-Hee, P., Chuan-Chin, P., and Wan-Young, C.
Nat. Softw. Engg. Conf. (NSEC), 1–7. (2014). Mobile-based kernel-fuzzy-c-means-wavelet
Yan, C., Frans, C., Yong, Y., Xiaosong, Y., and Bailing, Z. for driver fatigue prediction with cloud computing.
(2016). Video-based classification of driving behavior Sensors 2014 IEEE, 1236–1239.
using a hierarchical classification system with mul- Hostens, I. and Herman, R. (2005). Assessment of muscle
tiple features. Int. J. Pat. Recogn. Artif. Intel., 30(05), fatigue in low level monotonous task performance
1650010. during car driving. J. Electromyograp. Kinesiol., 15(3),
Teyeb, I., Olfa, J., Mourad, Z., and Chokri, B. A. (2014). 266–274.
A novel approach for drowsy driver detection using Katsis, C. D., Ntouvas, N. E., Bafas, C. G., and Fotiadis, D.
head posture estimation and eyes recognition system I. (2004). Assessment of muscle fatigue during driving
based on wavelet network. IISA 2014 5th Int. Conf. using surface EMG. In Proceedings of the IASTED in-
Inform. Intel. Sys. Appl., 379–384. ternational conference on biomedical engineering, vol.
Katyal, Y., Suhas, A., and Shipra, D. (2014). Safe driving by 262. doi: 10.2316/Journal.216.2004.2.417-112
detecting lane discipline and driver drowsiness. 2014 Balasubramanian, V. and Adalarasu, K. (2007). EMG-based
IEEE Int. Conf. Adv. Comm. Con. Comput. Technol., analysis of change in muscle activity during simulated
1008–1012. driving. J. Bodywork Mov. Ther., 11(2), 151–158.
Ingre, M., Torbjörn, Å., Björn, P., Anna, A., and Göran, K. Hu, S. and Gangtie, Z. (2009). Driver drowsiness detection
(2006). Subjective sleepiness, simulated driving per- with eyelid related parameters by support vector ma-
formance and blink duration: examining individual chine. Exp. Sys. Appl., 36(4), 7651–7658.
differences. J. Sleep Res., 15(1), 47–53. Lal, S. K. L. and Ashley, C. (2001). A critical review of the
Fairclough, S. H. and Graham, R. (1999). Impairment of driv- psychophysiology of driver fatigue. Biol. Psychol.,
ing performance caused by sleep deprivation or alcohol: 55(3), 173–194.
a comparative study. Human Factors, 41(1), 118–128. Khushaba, R. N., Sarath, K., Sara, L., and Gamini, D.
Thiffault, P. and Jacques, B. (2003). Monotony of road en- (2010). Driver drowsiness classification using fuzzy
vironment and driver fatigue: a simulator study. Acc. wavelet-packet-based feature-extraction algorithm.
Anal. Preven., 35(3), 381–391. IEEE Trans. Biomed. Engg., 58(1), 121–131.
Li, Z., Shengbo, E. L., Renjie, L., Bo, C., and Jinliang, S. Kukreja, V. and Sakshi. (2022). Machine learning models
(2017). Online detection of driver fatigue using steer- for mathematical symbol recognition: A stem to stern
ing wheel angles for real driving conditions. Sensors, literature analysis. Multimedia Tools Appl., 81(20),
17(3), 495. 28651–28687.
Zhenhai, G., Le, D., Hu, H., Yu, Z., and Wu, X. (2017). Driv- Chieh, T. C., Mohd, M. M., Aini, H., Seyed, F. H., and
er drowsiness detection based on time series analysis Burhanuddin, Y. M. (2005). Development of vehicle
of steering wheel angular velocity. 2017 9th Int. Conf. driver drowsiness detection system using electroocu-
Meas. Technol. Mech. Autom. (ICMTMA), 99–101. logram (EOG). 2005 1st Int. Conf. Comp. Comm. Sig.
Akin, M., Muhammed, B. K., Necmettin, S., and Muhittin, Proc. Special Track Biomed. Engg., 165–168.
B. (2008). Estimating vigilance level by using EEG and Sharma, R. and Vinay, K. (2022). Segmentation and multi-
EMG signals. Neural Comput. Appl., 17, 227–236. layer perceptron: An intelligent multi-classification
Lin, F.-C., Li-Wei, K., Chun-Hsiang, C., Tung-Ping, S., and model for sugarcane disease detection. 2022 Int. Conf.
Chin-Teng, L. (2012). Generalized EEG-based drowsi- Dec. Aid Sci. Appl. (DASA), 1265–1269.
20 Issues with existing solutions for grievance redressal
systems and mitigation approach using blockchain
network
Harish Kumar, Rajesh Kumar Kaushala and Naveen Kumar
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
Grievance redressal has always been vital for any organization to maintain a good work environment for its stakeholders.
Some organizations follow online portals, websites, or mobile applications to register grievances to provide more privacy to
the complainant’s identity. However, online platforms provide better solutions to the existing manual methods for grievance
redressal. Still, there are a lot of issues and challenges associated with them. This research has comprehensively analyzed the
existing grievance redressal systems to identify and discuss all the challenges. After comprehensive analysis, it is found that
presently there are several issues such as delayed response, opaque processes, biases, complexity and accessibility issues, lack
of personalization, and other privacy and security concerns associated with existing grievance redressal methods. To address
all these issues this study is proposing a blockchain-based solution for grievance redressal systems. The proposed solution
will be a blockchain-based web and mobile application that consists of multiple entities such as complainants, redressal
committee, and higher authorities. This system will provide the necessary privacy and confidentiality to the complainants
through the immutable distributed ledger technology and auditability of the entire process with complete transparency.
[Link]@[Link]
a
Applied Data Science and Smart Systems 141
Research objectives
The primary objective of this research article is to
investigate the challenges faced by existing grievance
redressal systems and explore how blockchain tech-
nology can mitigate these issues. To achieve this, the
following research goals will be pursued:
IV. Methodology
The research work is conducted using a prisma
approach in order to ensure its conclusion. A stan-
dardized methodology encompassing the stages of
planning, execution, and reporting was implemented.
The sections that follow outline the procedural phases
of the approach employed in this study. The initial
stage involves the formulation of search keywords.
The subsequent stage is conducting a search for sys-
tems designed to address grievances, (Figure 20.1)
online portals dedicated to grievance redressal, and
research papers pertaining to blockchain technology,
utilizing certain keywords. The purpose of utilizing
certain keywords is to effectively distinguish between Figure 20.2 Structure of grievances redressal system
research articles that are relevant and those that are
irrelevant. In the third phase, an analysis is conducted
on the operational characteristics of online portals. of a proficient and impactful method for addressing
This examination involves the identification of short- grievances is an essential requirement for any orga-
comings in current solutions and the subsequent nization or institution to demonstrate responsibility
proposal of blockchain technology as a fundamental and accountability (Tripathi, Srivastava, and Singh,
remedy for the aforementioned concerns within this 2021). The workflow of the traditional grievances
industry (Figure 20.2). redressal system is shown in Figure 20.2.
Grievances may emerge at several levels within an
organization, including educational institutions such
V. Related work
as universities and schools. When any individuals per-
A complaint is typically characterized as a form of ceive that their rights, needs, or expectations have not
communication, whether spoken or written that been adequately fulfilled. This issue becomes highly
articulates dissatisfaction with a particular course delicate when it pertains to the students of an aca-
of conduct or neglect, or with the quality of service demic institution, given that students are the most
provided by an organization. The implementation vulnerable individuals in this context. Frequently,
142 Issues with existing solutions for grievance redressal systems
individuals encounter difficulties in effectively com- have encompassed the crucial features which are
municating their concerns and encounter challenges required for an effective and efficient system, such as
in receiving adequate help from relevant authorities at immutability, transparency, auditability, and distrib-
different stages of their academic progression within uted storage to mitigate the risk of a single point of
the institution (Prajapat, Sabharwal, and Wadhwani, failure within the system.
2018). In a research article, the authors (Magner, Table 20.1 exhibits a comprehensive literature
1995) have examined a particular case wherein a sub- assessment of the existing solution for grievance
stantial group of students collectively endorsed and redressal on the basis of the key features of an effec-
submitted a petition alleging substandard teaching by tive and efficient system.
their teacher, citing an inability to effectively deliver The authors Prajapat, Sabharwal, and Wadhwani
the curriculum in accordance with the updated edu- (2018) in their study have proposed an automated
cational framework. The authors (Miklas and Kleiner, system for grievance registration and redressal, but it
2003) discussed the case of a foreign university where follows a very basic architecture and it is still human-
a group of female students registered a complaint dependent to forward the complaint at almost every
against their professor for harassment. stage, and due to that resolution to the student griev-
Many researchers proposed a variety of theoreti- ance may be delayed, the privacy of the user’s identity
cal frameworks, prototypes, online solutions, mobile is not preserved, security of sensitive data is not men-
applications, and web portals utilizing diverse tech- tioned and covered, due to centralized system archi-
nologies such as artificial intelligence and machine tecture, tampering with the data may be possible,
learning (ML) techniques to manage grievance and due to weak architecture it cannot be implemented
redressal processes. However, none of these proposals at larger scale. The authors Kandhari and Mohinani
(2014) have designed a mobile application for the citi- In the event of prolonged inactivity, the system auto-
zens to register their municipal services-related griev- matically elevates the complaint’s status to the supe-
ances. The application enables the users to register rior officer and District Magistrate through an email
their complaints along with the image of problems notification, providing an update on the registered
and location coordinates. The author Kormpho et al. complaint. The authors Shettigar et al. (2021) in a
(2018) proposed a solution that involves the devel- separate study proposed a blockchain-based solution
opment of a mobile application and a chatbot that for a grievances management system for college stu-
allows end-users to effectively register their concerns. dents, being a blockchain-based solution, it provides
Additionally, an innovative web-based solution is all the inherent features of blockchain but the most
provided for the organization to address these issues, important phase, which is implementation, is miss-
with the added benefit of preventing duplicate com- ing. The proposed solution uses the permissioned
plaints. The authors Palanissamy and Kesavamoorthy blockchain hyperledger fabric framework. In another
(2019) in their proposed solution comprise a widely study, the researchers Jha, Sonawane, and others
used online application for addressing grievances. (2022) presented a proposal for the development
This approach employs a multi-step negotiation pro- of a web portal designed specifically for students to
cess to identify and resolve issues. Which depends on register their complaints across different categories.
human involvement at each level. However, the pro- This portal offers transparency and monitoring capa-
posed method provides an easy user interface but is bilities, allowing the tracking of complaint statuses
deficient in key attributes such as tamper-proofing, at any given stage. Additionally, the authors suggest
immutability, privacy, and transparency, all of which the incorporation of ML and artificial intelligence
are of utmost significance when dealing with sensitive (AI) techniques to identify and address instances of
information. offensive language and complaints that propagate
In another study, the authors Aravindhan et al. misinformation on sensitive subjects such as racism,
(2020) developed a web-based solution that relies gender, and religion.
on human intervention at various stages to address In another study, authors Musa et al. (2021) have
complaints. However, this dependence on human proposed a centralized web portal to handle the stu-
involvement can potentially lead to delays in resolv- dents’ grievances at the university level, where stu-
ing student grievances. Furthermore, the preservation dents can register their academic and non-academic
of user identification and the protection of sensitive related grievances. Another govt. of India, initiative
data are not well-addressed or discussed. The poten- Rana et al. (2016) has introduced an online portal
tial for data tampering exists due to the central- for Indian citizens, where they can register their com-
ized system architecture, and the proposed solution plaints against any central or state govt. departments,
is not feasible for larger-scale implementation due sub-departments, or any public service-providing
to its inherent weaknesses in architecture. The sole agencies.
advantageous aspect of the suggested solution is to The Director of Public Grievances, The Department
the incorporation of a web interface into the preex- of Administrative Reforms and Public Grievances has
isting manual system. The researchers Laxmaiah and implemented a web-based portal to address and mon-
Mahesh (2020) of the study developed an automated itor public grievances. This portal is interconnected
and intelligent mobile application for the citizens to with all ministries and departments of the Indian gov-
register their grievances, the mobile application uses ernment as well as state governments. It allows citi-
ML techniques to segregate the types of complaints zens to access the portal through a mobile application
and automatically forward them to the concerned and register their grievances about any service pro-
department of an official without any delay, most of vided by the Indian government or state government.
the phases of this system is totally automated without Additionally, the portal offers transparency to users
any human dependency or intervention, researchers by enabling them to track the progress of their griev-
use cloud vision server and geo-coding and reverse ances using a unique registration number (DARPG,
geo-coding to label and identify the problem location 2023). The success of online portals for grievance
without any human involvement. It also provides the redressal systems has been comprehensively assessed
user’s accounts and their registered grievances and and analyzed by the authors Rana et al. (2015) in a
also provides the tracking information about the reg- separate study. This evaluation was conducted using
istered complaints. an E-government-based IS success model, which
In another study, the authors Hingorani et al. was constructed utilizing existing IS success models.
(2020) proposed a police complaint system by utiliz- Multiple hypotheses were examined and supported
ing blockchain technology in the development of web by empirical evidence, indicating that the implemen-
and mobile applications. It enables complainants to tation of the online public grievances redressal system
conveniently register and monitor their complaints. is likely to be highly effective. However, it is important
144 Issues with existing solutions for grievance redressal systems
Delayed resolutions
The effectiveness of conventional methods for
addressing grievances is often hampered by lengthy
and extended procedures for resolving disputes. It can
result in extended suffering for the aggrieved parties,
especially in cases where time-sensitive issues are at
stake. Delays can also increase tensions, and conflicts
which aggravate disputes, hence emphasizing the sig-
nificance of quick resolution.
Figure 20.3 Publication trends in existing solutions on
grievances redressal system
Susceptibility to manipulation
Several grievance redressal systems exhibit vulnerabil-
to note that despite the convenience and accessibility ity to manipulation, stemming from either unethical
offered by the online portal for registering grievances, practices or organizational inefficiency. This suscep-
the privacy and security of user data remain signifi- tibility undermines the justice of the system and may
cant concerns. discourage individuals from seeking resolution for
The researchers Alawneh, Al-Refai, and Batiha their issues early.
(2013) in their study investigated the factors that
influence user satisfaction with Jordan’s e-govern- Lack of accountability
ment services portal. The research paper outlines five Accountability counted as a key component of an effi-
primary criteria that have the potential to influence cient grievance redressal system. However, in several
the level of satisfaction among Jordanian individu- cases, it proves to be quite a challenge to ensure that
als with the portal. These factors encompass security the individuals or entities involved are held liable for
and privacy, trust, accessibility, awareness of public their actions. The absence of accountability can give
services, and the quality of public services. The study rise to a culture of freedom, wherein instances of mis-
presents significant findings derived from the analysis conduct remain unaddressed.
of survey data, emphasizing the importance of com-
prehending these factors to enhance the design and Inadequate data security
functionality of e-government portals. It also provides The rising dependence on online platforms for the
recommendations for practitioners and policy-mak- resolution of grievances has led to increased attention
ers to effectively improve user experience and cater on data security. The occurrence of breaches and data
to the needs of citizens. The outcomes of this study leaks can result in significant implications, such as the
underscore the shortcomings of the current system for disclosure of confidential data and a decline of confi-
addressing issues. dence in the system.
The literature review indicates that most of the
studies have implemented a centralized solution, few VII. Proposed system
articles only discuss the theoretical models, and few
have done the analysis of the effectiveness of the exist- The main purpose of this study is to investigate the
ing centralized solution, and the blockchain-based existing literature and identify the issues with the
studies are merely proposing a theoretical model. As existing solutions and propose an effective and effi-
per the reviewed literature, Figure 20.3 illustrates the cient system for grievance redressal which will cover
publication trends observed in published articles in all the issues identified during the literature review
this domain. of the existing solutions. The proposed system will
be an online web/mobile application that will use a
blockchain framework to provide distributed stor-
VI. Issues identified with existing grs age and provide immutability and full auditability.
Lack of transparency The structure of the proposed system is shown in
One of the significant challenges in existing grievance Figure 20.4.
redressal procedures is the absence of openness. In
several instances, people are often uninformed of the VIII. Result and discussion
status of their grievances, the procedures involved in
decision-making, and the eventual outcome. The lack The conventional methods of addressing grievances
of transparency can result in feelings of dissatisfaction, are considered to be inadequate due to their limited
Applied Data Science and Smart Systems 145
effectiveness in delivering crucial elements necessary IX. Conclusion and future work
for an efficient solution, such as transparency, immu-
Blockchain can provide transparency, immutability,
tability, privacy, quick redressal, security, and audit-
and a distributed storage facility and its integration
ability. The prevailing approach in online grievance
with various sectors can improve the existing services.
management solutions relies on a centrally managed
This article explores the limitations and issues of the
server system, rendering them more vulnerable to
present grievance redressal procedures. Although some
potential removal or tampering of data. Conversely,
of the online portals provide some sort of privacy and
the implementation of a decentralized grievance
transparency but they fail to provide immutability and
redressal system may hinder these efforts due to the
solution to a single point of failure. While identifying
widespread availability of all grievances across every
the limitations of the existing grievances redressal sys-
peer within the network.
tem this study proposes blockchain technology as a
During the literature review on the traditional and
mitigation approach to all the identified issues with
other online solutions for the grievances redressal sys-
the existing solutions. Blockchain-powered systems
tem, some important issues are identified. This exhib-
offer benefits like increased transparency, record pres-
its that there is a strong requirement for an efficient
ervation, and decentralized trust mechanisms.
grievance redressal system that should be both trans-
The findings of our investigation indicate that the
parent and tamper-proof and operate on a distributed
adoption of a blockchain-powered grievance redressal
peer-to-peer network which eliminates any potential
system presents numerous benefits, such as increased
instances of ignorance and abuse of power by higher-
transparency, the preservation of unalterable records,
level officials.
and the utilization of decentralized trust mechanisms.
Our proposed system extends the security, pri-
Through the utilization of smart contracts and cryp-
vacy, and other key features of the online portal by
tographic methodologies, blockchain technology has
adding blockchain technology, which provides all
the potential to enable a secure and effective mecha-
the inherent features of blockchain such as immu-
nism for addressing grievances. However, scalability,
tability, auditability, transparency, and distributed
privacy, and regulatory constraints are crucial and
ledger storage which mitigate the single point of
that can be managed with the selection of an appro-
failure of a centralized system and auditability fea-
priate blockchain framework.
ture enable the authorized user to check the com-
plete transaction history, and transparency feature
of blockchain allows the users to get updates about References
every change/transaction made to the registered Aggarwal, A., Ran Singh, D., and Kamrunnisha, N. (2018).
complaint. Impact of structural empowerment on organizational
146 Issues with existing solutions for grievance redressal systems
commitment: The mediating role of women’s psycho- Laxmaiah, M. and Mahesh, K. (2020). An intelligent public
logical empowerment. Vision, 22(3), 284–294. grievance reporting system-IReport. ICCCE 2020 Proc.
Alawneh, A., Al-Refai, H., and Batiha, K. (2013). Mea- 3rd Int. Conf. Comm. Cyber Phy. Engg., 197–207.
suring user satisfaction from e-government services: Magner, D. K. (1995). Mid-semester removal of Professor
Lessons from Jordan. Gov. Inform. Quart., 30(3), Roils University of Montana. Chron. Higher Edu., A25.
277–288. Miklas, E. J. and Brian, H. K. (2003). New developments
Aravindhan, K., Periyakaruppan, K., Aswini, K., Vaishnavi, concerning academic grievances. Manag. Res. News,
S., and Yamini, L. (2020). Web portal for effective stu- 26(2/3/4), 141–147.
dent grievance support system. 2020 6th Int. Conf. Musa, W. M. W., Azahari, A. A., Noorimah, M., and Suhaily,
Adv. Comput. Comm. Sys. (ICACCS), 1463–1465. M. A. M. (2021). E-justice: Students’ complaints made
Bhadouria, L. S. and others. (2021). Online complaint man- easy. SEARCH J. Media Comm. Res. (SEARCH), 1.
agement system. Turkish J. Comp. Math. Edu. (TUR- Oguntosin, V., Oluwadurotimi, M., Adoghe, A., Abdulka-
COMAT), 12(6), 5144–5150. reem, A., and Adeyemi, G. (2021). Development of a
DARPG. (2023). Centralized public grievance redress and web-based complaint management platform for a Uni-
monitoring system. [Link] versity community. J. Engg. Sci. Technol. Rev., 14(1),
AboutUs. 150–159.
Denny, J., Ramya, C., Sweta, R. L., Srija Reddy, A., and Sa- Palanissamy, A. and Kesavamoorthy, R. (2019). Automated
hithya, V. (2021). A Web Portal for Student Grievance dispute resolution system (ADRS) – A proposed ini-
Support System. International Research Journal of tial framework for digital justice in online consumer
Engineering and Technology. 8(5), 1261–1263. transactions in India. Proc. Comp. Sci., 165, 224–231.
Hingorani, I., Rushabh, K., Deepika, P., and Nataasha, R. Prajapat, S., Vaibhav, S., and Varun, W. (2018). A prototype
(2020). Police complaint management system using for grievance redressal system. Proc. Int. Conf. Recent
blockchain technology. 2020 3rd Int. Conf. Intel. Sus- Adv. Comp. Comm. ICRAC 2017, 41–49.
tain. Sys. (ICISS), 1214–1219. Rana, N. P., Yogesh, K. D., Michael, D. W., and Banita, L.
Jattan, S., Vineeth, K., Akhilesh, R., Rachith, R. N., and Sne- (2015). Examining the success of the online public
ha, N. S. (2020). Smart complaint redressal system us- grievance redressal systems: An extension of the IS
ing Ethereum blockchain. 2020 IEEE Int. Conf. Dis- success model. Inform. Sys. Manag., 32(1), 39–59.
tribut. Comput. VLSI Elec. Cir. Robot. (DISCOVER), Rana, N. P., Yogesh, K. D., Michael, D. W., and Vishanth, W.
224–229. (2016). Adoption of online public grievance redressal
Jha, S., Pankaj, S., and others. (2022). Smart student system in India: Toward developing a unified view.
grievance redressal system with foul language detec- Comp. Human Behav., 59, 265–282.
tion. 2022 8th Int. Conf. Adv. Comp. Comm. Sys. Shahnawaz, M., Prashant, S., Prabhat, K., and Anuradha, K.
(ICACCS), 1, 187–192. (2020). Grievance redressal system. Int. J. Data Min.
Kandhari, V. K. and Keertika, D. M. (2014). GPS based com- Big Data, 1(1), 1–4.
plaint redressal system. 2014 IEEE Global Human. Shettigar, R., Nishant, D., Ketan, I., Farhan, A., and Ram-
Technol. Conf. South Asia Satel. (GHTC-SAS), 51–56. krushna, C. M. (2021). Blockchain-based grievance
Kormpho, P., Panida, L., Narut, P., and Siripen, P. (2018). management system. Evol. Computat. Intel. Front.
Smart complaint management system. 2018 Seventh Intel. Comput. Theory Appl. (FICTA 2020), 1,
ICT Int. Student Project Conf. (ICT-ISPC), 1–6. 211–222.
Kumar, A., Sharad, S., Nitin, G., Aman, S., Xiaochun, C., Tripathi, U. N., Amit Kumar, S., and Bhanu P. S. (2021). Ef-
and Parminder, S. (2021). Secure and energy-efficient fectiveness of online grievance redressal and manage-
smart building architecture with emerging technology ment system: A case study of IGNOU learners. Ind. J.
IoT. Comp. Comm., 176, 207–217. Edu. Technol., 3(2), 92.
21 A systematic approach to implement hyperledger fabric
for remote patient monitoring
Shilpi Garg, Rajesh Kumar Kaushala and Naveen Kumar
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
The integration of blockchain technology, particularly hyperledger fabric, into the domain of remote patient monitor-
ing, presents a new era that has the potential to greatly improve healthcare systems. This research paper introduces a
methodical strategy for integrating hyperledger fabric into remote patient monitoring systems. It provides a structure that
efficiently addresses key challenges pertaining to data security, privacy, and interoperability. This study aims to establish a
methodology for remote patient monitoring environments by carefully analyzing the distinctive requirements and constraints
associated with these types of environments. The methodology encompasses various key phases, including network configu-
ration, smart contract design and sensitive data management specifically tailored to healthcare contexts. In addition, the
study explores practical methods of implementation and conducts performance evaluation of the suggested strategy using
minifab and hyperledger explorer, respectively. This analysis provides valuable insights into the effectiveness and efficiency
of the approach in safeguarding the privacy and security of patient information. Through an examination of the mutually
beneficial capabilities of hyperledger fabric and remote patient monitoring, this study makes a valuable contribution to the
advancement of healthcare systems that are both secure and efficient.
Keywords: Hyperledger fabric, remote patient monitoring, blockchain, hyperledger explorer, minifab
a
[Link]@[Link]
148 A systematic approach to implement hyperledger fabric for remote patient monitoring
et al., 2020; Tanwar, Parekh, and Evans, 2020). The sibility of transmitting information among net-
technology was developed within the framework of work participants while upholding data integrity.
the “Hyperledger Foundation”, an organization led Additionally, it enables the establishment of spe-
by IBM. It possesses several notable features, such as cific criteria or permissions to encapsulate the
the ability to create private data collections, strong transmitted data. In situations where maintain-
security measures for Docker containers, a flexible ing the confidentiality of particular information
programming framework, and a consensus model that is crucial, the option exists to create a separate
can be adjusted based on the host nodes. Hyperledger channel distinct from the rest, accessible only
fabric consists of various major components includ- to select organizations. This feature underscores
ing peers, orderer, chaincode, membership service the potential for multiple blockchains to coexist
provider (MSP), channels, and fabric certificate within the same network, as a channel essentially
authority (CA). Figure 21.1 illustrates the transaction operates as an independent blockchain.
flow diagram of the hyperledger fabric (Pongnumkul, • Certification authorities (CAs) are a fundamental
Siripanpornchana, and Thajchayapong, 2017; component of public key infrastructures (PKIs)
Performance, Group, and others, 2018; Jennath, and have been assigned with ensuring the distri-
Anoop, and Asharaf, 2020; Woznica and Kedziora, bution of digital certificates. The primary func-
2022). tion of this layer is to verify the identities of the
Peer refers to the individual nodes that comprise parties or actors involved in the communication,
the network organizations. The aforementioned ensuring that they are indeed who they claim to
pieces are responsible for providing information to be. Websites commonly possess a digital certifi-
the ordering nodes within the network, enabling them cate that is issued by a reputable CA in order to
to configure the blocks that are being transacted. authenticate the trustworthiness of the visited
Orderer: One of the pivotal components within the website.
network, the orderer assumes a critical role in config- • The membership service provider (MSP) is re-
uring blocks according to specified criteria and dis- sponsible for gathering all cryptographic tech-
tributing them to their respective peers. These peers niques employed for network interaction. It is
can be affiliated with one or multiple organizations, imperative for every organization to own a man-
necessitating the attainment of a consensus agreement aged security provider (MSP) that encompasses
as per the network’s requirements. All transactions its cryptographic data, including keys and the CA
related to network configuration flow through the responsible for issuing its certificates. The cre-
orderer. Additionally, these computing entities enforce dentials are utilized by clients for the purpose of
fundamental access control for channels, determining authenticating their transactions, while peers em-
who has read and write privileges and the authority ploy them to authenticate the outcomes of trans-
to configure them. action processing, specifically endorsements.
• Chaincode, often referred to as smart contracts
• Channel functions as a communication medium within the context of hyperledger fabric, serves
among network participants. In this context, it as the mechanism through which contractual
serves as a mechanism for conducting private agreements are implemented. A smart contract
communications, ensuring data isolation and refers to a block of code that is triggered by an
confidentiality. This layer takes on the respon- external client application, operating outside
Applied Data Science and Smart Systems 149
the blockchain network. Its purpose is to over- comprising a network of peers. Language java is uti-
see the manipulation and control of a collection lized for the purpose of writing chaincode. Figure 21.3
of key-value pairs inside the present state of the is a screenshot of [Link] file that is a configuration
network, accomplished through the execution of file about the network used by minifab. Table 21.1
transactions. Smart contracts are encapsulated
and distributed as chaincode. Subsequently, the
chaincode is deployed onto the peers and subse-
quently defined and utilized within one or many
channels.
III. Implementation
In order to effectively handle the patient data it is
necessary to establish a correlation between the vari-
ous components of the fabric and the demands of the
RPM-based EHR systems. All medical centers func-
tion as entities inside a fabric network. The patient
data has been regarded as valuable resources stored
within the ledger. Currently, patient records consist
of a limited number of categories, encompassing
personal and medical information such as age, resi- Figure 21.3 [Link] file for network
dence, allergies, symptoms, therapy, follow-up, and
so on. When a physician administers medication to a
Table 21.1 Process to build up the minifab network.
patient, they will have access to the patient’s medical
history data, which assists them in determining the Steps Command Description
most suitable type of medical care. Figure 21.2 illus-
1 minifab netup -s Start the network
trate the system architecture. couchdb -e true -i by adding hospital1.
The medical database is utilized to establish a 2.4.8 -o hospital1. [Link] as a
repository of transactions within the proposed sys- [Link] current organization
tem. The orderer and peer nodes are executed within 2 minifab create -c Create the
the Docker container. Hyperledger fabric framework healthchannel channel named as
is designed to be configured with a minimum of two healthchannel
organizations-hospital1 and hospital2. Every organi- 3 minifab join -c Network will join the
zation will be assigned to a single peer node, a chan- healthchannel healthchannel
nel, and an orderer node within the ordering service. 4 minifab Update the anchor
Each peer node within the network possesses a dupli- anchorupdate peer node
cate of the ledger. A chaincode is developed with the 5 minifab profilegen -c Generate the profiles
purpose of facilitating access to the two organizations healthchannel for healthchannel
Abstract
Kannada is one of the major regional languages of Karnataka, a prominent state of India. The text processing tasks are very
important and highly required for the development of the language in this digital world. Spell checking is one of the needs
in creating an effective document. Even though one can find several tools on the internet, it allows you to type or paste the
Kannada text on the text editor and submit the text then the result will appear on the another editor. The proposed work de-
lineates on developing an efficient interactive spell checking and transliteration tools for the Kannada language based on the
Blooms filter algorithm, Symspell technique and International Phonetic Alphabet (IPA) representation. This work carried out
with an intention to provide handy text processing tools to the public. The proposed work has been tested on several datasets
and found to be useful with more than 85% and 87.29% accuracy for both spell check and transliteration tools, respectively.
Keywords: Blooms filter, Symspell algorithm, Levenshtein distance, transliteration, International Phonetic Alphabet, candi-
date words
chandrika@[Link]
a
Applied Data Science and Smart Systems 153
language, regardless of the writing system used. The more effort into creating a real-word spell checker
IPA includes symbols for consonants, vowels, and that incorporates additional language principles.
other sounds, as well as diacritic marks that indicate A spell check tool using Levenshtein’s edit distance
variations in pronunciation, such as stress and tone. algorithm, rule-based algorithm, Soundex algorithm,
While IPA-based transliteration can be more precise and LSTM (Long- Short-Term Memory) model is
than other methods, it can also be more complex developed for Tamil language (Sampath et al., 2022).
and time-consuming, especially for those who are The model handles three categories of errors with a
not familiar with the IPA. Additionally, not all lan- good performance of 95.67%.
guages have a one-to-one correspondence between A Telugu spell-checker’s innovative concept and
their sounds and IPA symbols, which can lead to some implementation are presented in (Parameshwari et
ambiguity in transliteration. Despite these challenges, al., 2012). The core of Telugu spell-checking is mor-
IPA-based transliteration remains a valuable tool for phological validation using a morphological analyzer.
those who need to work with multiple languages and Along with issues affecting orthography and morphol-
writing systems. ogy, difficulties associated with Telugu document spell
In the proposed work, both dictionary and translit- checking are examined. On these lines, a spell-checker
eration tools are developed, the implementation part has been created. The spelling checker’s architecture
will focus more on these two models. and algorithm, which is based on Sandhi splitter and
morphological analysis principles, are described.
Related work Additionally, it contains tables of spelling variations
gleaned from Telugu’s spatiotemporal dialects.
A comprehensive survey is done on various languages A common approach used to develop a spell check
to understand the methodology/ technique used by tool is minimum edit distance algorithm (Patil et al.,
researchers. 2021). By carrying out numerous operations including
The researchers have explored methodologies character replacement, insertion, and deletion, it fixes
(Randhawa et al., 2014) used for developing spell spelling mistakes. The proposed work focuses on cor-
check tool for various Indian regional languages recting the errors for Marathi text and it works better
including the performance analysis. This helps a for short words with a good accuracy of 85.5%.
researcher to understand the pros and cons of avail- The challenges associated with multi-lingual
able techniques. A spell checker tool on Bangla is speech recognition and propose solutions to address
explored in (Chaudhuri et al., 2002). Researchers these challenges in the Indian context are explored in
have handled the errors based on the phonetic in two (Manjunath et al., 2019; Khattar et al., 2020). The
stages using phonetically similar character error cor- proposed model contributes to the advancement of
rection and reversed word dictionary and error cor- speech technology in the context of Indian languages,
rection. Experiment is conducted on three million which is crucial for enabling effective communication
words which are arranged in Trie data structure and and technology access for the diverse linguistic popu-
obtained satisfied results. lation in India.
Speech recognition is explored in (Priya et al., Both forward and backward transliteration of
2022), authors have used novel Automatic Speech Punjabi names was performed between Gurmukhi
Recognition system for seven low-resource languages and English Roman scripts using an n-gram language
based on deep sequence modeling with an enhanced model (Goyal et al., 2022). Over one million paral-
spell checker. The researchers have obtained word lel entities of person names in both scripts were used
error rate (WER) of 0.62 using recurrent neural as the training corpus. The study created extensive
network-gated recurrent unit (RNN-GRU) and the English-to-Punjabi and Punjabi-to-English n-gram
transformer-based INDIC Bidirectional Encoder pre- databases, comprising more than 10 million n-grams
sentations significantly enhance performance by 10% with multiple script mappings. Categorizing n-grams
and lower the average WER to 0.52. into starting, middle, and ending n-grams was essen-
A spell check tool is explored on Dawurootsuwa tial due to variations in pronunciation based on let-
which is one of the Ethiopian languages (Arya et al., ter placement in words. The transliteration process
2021; Gamu et al., 2023), it has poor dataset. The involved searching for the longest matching n-gram
root words in this study were built using the Hunspell in the database, recursively splitting the string until a
dictionary format and consisted of 5,000 total root match was found, and then merging the transliterated
words, more than 2,500 morphological rules, and strings to produce the final output.
3,156 unique terms for testing. total spell error detec- The challenges of speech recognition and spell cor-
tion performance was 90.4%, and total spell error rection in low-resource Indian language are discussed
repair performance was 79.31%, according to the in the work did by Priya et al. (2022). The authors
experimental results. Additionally, we are putting propose a solution using Indic BERT. A multi-lingual
154 Developing spell check and transliteration tools for Indian regional language – Kannada
transformer-based language model, the model per- regular keyboard (English language keyboard)
forms speech recognition and spell correction for the and can check the correct Kannada words on the
text written in Tamil, Telugu, and Kannada languages. editor.
By leveraging the power of transfer learning, the
authors demonstrate the effectiveness of Indic BERT Methodology
in improving the accuracy of speech recognition and
spell correction tasks in these languages. The general steps to develop a spell-checking tool in
Kannada speech corpus for automatic speech rec- Kannada language is as follows:
ognition system based on phoneme (Praveen et al.,
2022) is developed for Kannada corpus. The authors • Corpus collection: Gather a large collection of
describe the methodology employed in creating the correctly spelled Kannada text. This can include
corpus, which includes collecting speech samples books, articles, websites, and other reliable sourc-
from native Kannada speakers and annotating them es written in Kannada.
with phoneme-level transcriptions. The resulting • Corpus pre-processing: Clean and preprocess the
corpus serves as a valuable resource for researchers collected corpus data by removing any unwanted
and practitioners working on Kannada speech recog- characters, punctuation marks, and special sym-
nition, enabling the development and evaluation of bols. Normalize the text to ensure consistent rep-
accurate and efficient speech recognition models for resentation.
this language. • Tokenization: Split the pre-processed text into in-
A convolutional neural network-based speech rec- dividual words or tokens. This step helps in ana-
ognition model for Kannada Language is demon- lyzing and processing each word separately.
strated in work by Rudregowda et al. (2020). The • Build a dictionary: Create a dictionary of correct-
authors propose a methodology for visual speech ly spelled Kannada words based on the tokenized
recognition in Kannada. The findings of this study corpus. This dictionary will serve as the reference
contribute to the advancement of speech recognition for spell-checking.
technology for Kannada, which could have signifi- • Error generation: Generate a set of common
cant implications for speech-based applications in the spelling errors that occur in Kannada. This can
Kannada-speaking community. include typos, phonetic errors, and other com-
mon mistakes made by Kannada speakers.
• Spell-checking algorithm: Implement a spell-
Scope of the work
checking algorithm that compares each word in
From the survey, it is found extensive research work the input text with the words in the dictionary.
has not carried out in this domain. There is a lot of The algorithm should identify potential spelling
scope in this area. Summary of the survey is as fol- errors and suggest corrections based on the clos-
lows: After analyzing the survey, it is found that est matching words in the dictionary.
• User interface: Develop a user-friendly inter-
• In Kannada languages, minimum work has been face where users can input their text for spell-
carried out in spell check and transliteration do- checking and receive suggestions for correc-
main. tions. This can be a web-based interface or an
• Getting the proper Kannada datasets for training application.
and testing is not an easy task. • Testing and refinement: Test the spell-checking
• There is no open-source optical recognition tool tool with a variety of Kannada texts, including
available to convert pdf to word which is required different genres and writing styles. Collect user
for the text processing. feedback and refine the algorithm and user inter-
face based on the feedback received.
Objectives • Continuous improvement: Maintain and update
the spell-checking tool by periodically updating
From the survey, it is noted that, in Kannada language the dictionary with new words and refining the
there is a lot of scope with respect to transliteration error generation algorithms to improve accuracy
and not many research articles are published. We have and coverage.
contributed in this domain by
It is worth noting that building a robust and
• Developing a spell check tool with the possible accurate spell-checking tool requires a considerable
features amount of linguistic expertise and computational
• Designing a transliteration tool for Kannada lan- resources. Collaborating with Kannada language
guage, where the user can type Kannada using experts or researchers in natural language processing
Applied Data Science and Smart Systems 155
(NLP) would be beneficial in ensuring the effective- • Initialize a bit array of the specified size and set
ness of the tool. all bits to 0.
• Calculate the optimal number of hash functions
Dictionary-based spell checking tool based on the desired false positive probability
and the size of the dataset.
A huge dataset of 7 lakh is collected and in that • Create a list of hash functions using different seed
125,000 words are identified as unique words. These values.
words can have spelled in many ways all those mis-
spelled forms of these unique words are tabulated in a The following Figure 22.1 shows the architecture
dictionary which is used for error correction of the proposed model:
In this proposed model a user interface is devel-
oped such that it accepts the Kannada document or • Insert elements into the Bloom filter.
an editor is provided for the user to start typing the • For each element in the Kannada dataset, apply
Kannada articles. each hash function to generate hash values.
After the user uploads the document, two possible • Set the corresponding bits in the Bloom filter’s bit
scenarios can unfold. Firstly, a routine can be imple- array to 1 for each generated hash value.
mented to exhibit the precise contents of the docu- • Search for an element in the Bloom filter
ment within the designated text area. Secondly, all • Given a query element, apply each hash function
the words present in the document are divided into to generate hash values.
tokens, and the unique tokens are subsequently sub- • Check if the corresponding bits in the Bloom filter’s
jected to processing by Blooms filter. bit array are set to 1 for each generated hash value.
Here are the steps for implementing a Bloom filter • If any of the bits are not set to 1, the element is
searching algorithm for a Kannada dataset: definitely not present in the dataset.
• If all bits are set to 1, the element is probably pres-
• Create a Bloom filter: ent in the dataset (there is a false positive prob-
• Specify the desired size of the Bloom filter and the ability). Figures 22.2 and 22.3 shows the result of
number of hash functions to use. the search operation using Bloom filter.
deletions required to transform word1 into the given dataset. Similarly, the term fail in the graph
an empty string. refers to the percentage of failure in predicting the
7. Calculate the Levenshtein distance wrong words.
• Iterate through the characters of word1 From the Table 22.2 and the graph in Figure 22.5,
(from i=1 to m) and word2 (from j=1 to n). it is clear that for a small dataset like 10 words it
• If word1[i-1] is equal to word2[j-1] (i.e., the works pretty well with 90% accuracy. As we increase
characters are the same), the cost of the cur- the dataset it performs better, for 1 lakh of words
rent operation is 0. Set dp[i][j] to the value the accuracy is still better with 87%. Frequency of
of dp[i-1][j-1]. the words in the dictionary and different forms of
• If word1[i-1] is different from word2[j-1], grammatical words for a given word has an impact
we have three possible operations: on the performance of the model. If the data-
• Insertion: Calculate the cost of inserting set has more wrong words, then it will learn and
word2[j-1] into word1 at position i. Set dp[i] perform the prediction better. Figures 22.6–22.9
[j] to dp[i][j-1] + 1.
• Deletion: Calculate the cost of deleting
word1[i-1] from word1. Set dp[i][j] to dp[i- Table 22.1 Dataset details.
1][j] + 1.
• Substitution: Calculate the cost of substitut- Dataset Files
ing word1[i-1] with word2[j-1]. Set dp[i][j]
Articles 1026
to dp[i-1][j-1] + 1.
• Choose the minimum cost among the three Stories 51
possible operations and assign it to dp[i][j]. Wikipedia dataset 201
• Output: Grammar data 3
• The final Levenshtein distance is stored in
Dataset Content Size
dp[m][n], representing the minimum num-
ber of edits required to transform word1 Words 726,654
into word2.
Unique Words 179,863
The following Figure 22.4 demonstrates the results of
searching a word in a dictionary using Bloom filter.
Table 22.2 Performance analysis.
Results [Link]. Number of words Accuracy in %
The proposed work with complete user interface is
1 10 90
uploaded on a website and released to the public. The
website is designed by taking requirements from the 2 100 91
users working from various domains. Initially the 3 1000 89
model is tested with 7 lakh words and later with dif- 4 10,000 85
ferent set of words. Tables 22.1 and 22.2 shows the 5 100,000 87
dataset type, size and the accuracy of the model. The
graph in Figure 22.5 shows the performance of the
model. In the graph, the term pass refers to the accu-
racy of the model in predicting the wrong words for
Snapshots
Train 75,557
Validation 25,185
Test 25,185
Figure 22.7 Identifying wrong words (underlined in
red)
English to IPA translation
The model is built using IPA to transliterate words
written in English to Kannada language. The dataset
used has 125,927 unique words. It is represented as
each English word and all its IPA translations. The
Figure 22.10 demonstrates the abstract view of this
work.
The following are the steps involved in the
translation:
Data pre-processing
The input data is initially in the raw state, converting
the dataset into a pair of English words and their cor-
responding IPA translations is done in the preprocess-
ing stage by removing unwanted text like numbers
Figure 22.8 Selecting the wrong words with options
and the special symbols since these do not require any
translations.
The English words and IPA translations are
Kannada transliteration tool tokenized, and the unique characters in both sets are
This section describes the implementation details of extracted. The input sequences are padded to a fixed
transliteration which translates text from English to length to ensure uniformity. The dataset details are
Kannada. shown in the Table 22.3.
Applied Data Science and Smart Systems 159
Model architecture
Character BERT is a specialized variant of the BERT
model that operates at the character level, making
it suitable for tasks such as phonetic transcription.
When translating English words to IPA transcrip-
tions, character BERT learns the relationship between
input characters and their corresponding IPA sym-
bols. The process begins by encoding each English
word into individual characters and converting them
into numerical representations using a character
vocabulary.
The model architecture of character BERT con-
sists of a multi-layer bidirectional transformer that
captures contextual information from both the left
and right contexts of each character. Prior to fine-
tuning, Character BERT undergoes pre-training on
large-scale unlabeled data, where it learns to predict
masked characters based on the context provided by
surrounding characters. Figure 22.11 Character BERT embedding
During fine-tuning, the model is trained on a par-
allel dataset of English words and their IPA tran-
scriptions, enabling it to encode the input characters
and predict the correct IPA transcriptions using the
learned contextual information. In inference, given an
English word, the characters are tokenized, encoded,
and passed through the fine-tuned character BERT
model. The model generates a sequence of numeri-
cal representations that can be decoded using the IPA
vocabulary, yielding the corresponding IPA transcrip-
tion. Character BERT’s strength lies in its ability to
capture fine-grained information from individual
characters, enabling accurate and context-aware IPA
transcriptions for English words (Figure 22.11).
Model training
• The model is trained using the compiled model
with the RMSprop optimizer and categorical Figure 22.12(b) Mapping consonants to its IPA rep-
cross-entropy loss function. resentation
• The training is performed by providing the en-
coder input (English word sequence) and decoder
input (IPA translation sequence) to predict the de- • The generated IPA translations are outputted for
coder output (next IPA character) as shown in the evaluation or further processing.
Figures 22.12a and b.
• The model is trained on a training set and vali- After the IPA translation is generated, using the
dated on a separate validation set. IPA-Kannada mapping we map each IPA symbol
160 Developing spell check and transliteration tools for Indian regional language – Kannada
to a corresponding Kannada alphabet as shown in Table 22.4 Sample transliterations of Kannada words
the Figure 22.13. The pseudo code is given by the
following: English word IPA translation Transliteration
Conclusion
The proposed work incorporates both dictionary
based spell checking tool and transliteration. Spell
checking tool is based on dictionary and the per-
formance of the model mainly based on the volume
of the dictionary. As long as the dictionary is grow-
ing the performance starts improving. This is not an
effective nature instead if the model understands the
rules of the grammar, the model doesn’t depend on
the words in the dictionary. This is the planned work
in the future.
Kannada transliteration tool is a productive one
since it is based on the rules of the Kannada gram-
mar still more feature can be added to it in future and
released for public use.
References
Randhawa, Er, Sumreet, K., and Saroa, Er C. S. (2014).
Study of spell checking techniques and available spell
checkers in regional languages: a survey. Int. J. Tech-
Figure 22.13 Mapping consonants to its IPA nol. Res. Engg., 2(3), 148–151.
Applied Data Science and Smart Systems 161
Chaudhuri, B. B. (2002). Towards Indian language spell- Goyal, K. D., Muhammad, R. A., Vishal, G., and Yasir, S.
checker design. Lang. Engg. Conf., 2002 Proc., 139– (2022). Forward-backward transliteration of Punjabi
146. Gurmukhi script using n-gram language model. ACM
Priya, M. C., Shunmuga, D., Karthika, R., Ashok Kumar, Trans. Asian Low-Res. Lang. Inform. Proc., 22(2),
L., and Lovelyn Rose, S. (2022). Multilingual low re- 1–24.
source Indian language speech recognition and spell Manjunath, K. E., Dinesh Babu, J., Sreenivasa Rao, K., and
correction using Indic BERT. Sa-dhana-, 47(4), 227. Ramasubramanian, V. (2019). Development and anal-
Gamu, D. T. and Michael, M. W. (2023). Morphology-based ysis of multilingual phone recognition systems using
spell checker for Dawurootsuwa language. Scientif. Indian languages. Int. J. Speech Technol., 22, 157–168.
Prog., 2023. Arya, R., Singh, J., and Kumar, A. (2021). A survey of
Sampath, A. and Varadhaganapathy, S. (2023). Hybrid multidisciplinary domains contributing to affective
Tamil spell checker with combined character splitting. computing. Comp. Sci. Rev., 40, 100399, [Link]
Concur. Computat. Prac. Exp., 35(1), e7440. org/10.1016/[Link].2021.100399.
Patil, K. T., Bhavsar, R. P., and Pawar, B. V. (2021). Spell- Prakash, A. and Hema, A. M. (2022). Exploring the role of
ing checking and error corrector system for Marathi language families for building Indic speech synthesis-
language text using minimum edit distance algorithm. ers. IEEE/ACM Trans. Audio Speech Lang. Proc., 31,
Adv. Comput. Data Sci. 5th Int. Conf. ICACDS 2021, 734–747.
Nashik, India, April 23–24, 2021, Revised Selected Priya, M. C., Shunmuga, D., Karthika, R., Ashok Kumar,
Papers, Part I 5, 102–111. Springer International Pub- L., and Lovelyn Rose, S. (2022). Multilingual low re-
lishing, 2021. source Indian language speech recognition and spell
Priyadarshani, H. S., Rajapaksha, M. D. W., Ranasing- correction using Indic BERT. Sa-dhana-, 47(4), 227.
he, M. M. S. P., Kengatharaiyer, S., and Dias, G. V. Praveen, N. and Shashidhar, K. (2022). Phoneme based
(2019). Statistical machine learning for translitera- Kannada speech corpus for automatic speech recogni-
tion: Transliterating names between Sinhala, Tamil tion system. 2022 IEEE Int. Conf. Distribut. Comput.
and English. 2019 Int Conf. Asian Lang. Proc. Elec. Cir. Elec. (ICDCECE), 1–5.
(IALP), 244–249. Rudregowda, S., Sudarshan Patil, K., Gururaj, H. L., Vi-
Khattar, N., Singh, J., and Sidhu, J. (2020). An energy effi- nayakumar, R., and Moez, K. (2023). Visual speech
cient and adaptive threshold VM consolidation frame- recognition for Kannada language using VGG16 con-
work for cloud environment. Wire. Per. Comm., 113, volutional neural network. Acoustics, 5(1), 343–353.
349–367.
23 Real-time identification of traffic actors using YOLOv7
Pavan Kumar Polagania, Lakshmi Priyanka Siddib and Vani Pujitha M.c
Velagapudi Ramakrishna Siddhartha Engineering College, Vijayawada, India
Abstract
Real-time traffic object detection is a key topic in computer vision, especially for improving traffic safety and management. This
research describes a novel strategy for detecting traffic actors in real-time using YOLOv7, a cutting-edge deep learning system.
Traditional computer vision algorithms, such as Single Shot Detector, R-CNN, and older versions of You Only Look Once
(YOLO), frequently exhibit slow response times and poor accuracy in high-traffic areas.YOLOv7, an advanced object detec-
tion method based on convolutional neural networks (CNNs), is used in the proposed approach to address these difficulties
straight on. YOLOv7 not only achieves real-time object detection, but also greatly increases accuracy by removing superfluous
candidate boxes and employing a non-maximum suppression module to choose the best bounding boxes from overlapping
ones. Furthermore, the spatial pyramid pooling block improves accuracy by enhancing the network’s receptive field without in-
troducing additional parameters. In this study, we demonstrate the performance of our model under various driving scenarios,
including clear and cloudy skies, varying lighting, occlusions, and noisy input data. This model detects traffic participants such
as automobiles, pedestrians, cyclists, and traffic signs, which contributes to improved traffic safety and management.
offering distinct strengths and weaknesses, thereby Advantage: Enhanced intention recognition through
enriching the diverse landscape of solutions within human skeletal characteristics.
this domain. Disadvantage: Increased computational complex-
The method described in this study relies around the ity and longer training times due to multiple model
use of the Fast-Yolo-Rec method, which expertly bal- usage.
ances accuracy and speed. Its key goals are trajectory
classification via long- short-term memory (LSTM)- DFF-Net, which was introduced in this study, is
based recurrent networks and position prediction intended to detect real-world traffic items on rail-
via SSAM-YOLO and LSSN. The optical flow-based ways. It is divided into two parts: previous detection
detection method is critical in establishing the direc- and object detection. To initialize the system and
tion and speed of individual pixels inside a picture. restrict the search space for object detection, the pre-
An interesting method is used to speed up process- vious detection module employs VGG-16 pre-trained
ing. Odd frames of input images are designated for on ImageNet. The object detection module seeks to
detection, while even frames are committed to predic- recognize and predict the kinds of objects contained
tion, considerably increasing overall speed (Zarei et within the prior boxes (Li et al., 2020).
al., 2022).
Advantage: DFF-Net excels at increasing detection
Advantage: Fast-Yolo-Rec excels in rapid and cost- accuracy and effectively addressing class imbalance in
effective vehicle detection. railway object detection.
Disadvantage: However, it demands substantial com- Disadvantage: However, when compared to YOLO, a
putational resources for handling real-time data. one stage object detector, DFF-Net has a slower total
In this study, methodology introduces the SEF-Net speed.
framework, which is made up of three modules.
Stable bottom feature extraction (SBM), Lightweight The authors obtained a large dataset spanning
feature extraction (LFM), and Enhanced adaptive fea- numerous traffic incidents such as accidents, conges-
ture fusion module (EAM). SBM improves precision tion, and vehicle breakdowns in this study. They used
in tiny object detection by expanding convolutional a pre-trained Mask-SpyNet model for video-based
channels, which is especially beneficial for small object detection and post-processing to identify and
objects. Furthermore, attention enhancement blocks categorize traffic occurrences (Ye et al., 2021).
encode geographic and channel-specific semantic
Advantage: This novel approach considerably
information, which improves item detection and
enhances nighttime traffic event identification, hence
placement (Ye et al., 2022).
improving motorway traffic management safety and
Advantage: (1) This approach swiftly identifies car efficiency.
locations at a lower computational cost compared
Disadvantage: However, there are evaluation con-
to other high-speed detectors without necessitating
straints, and the method’s performance may be altered
additional processing. (2) SBM significantly enhances
by changing lighting circumstances.
precision for small object detection, outperforming
YOLOv4 in multi-detection capability. DLT-Net, the suggested technique in this study, is a
Disadvantage: Handling and analyzing large volumes unified neural network built for self-driving cars. Using
of real-time data demand substantial computational common features, it detects drivable zones, lane lines,
resources. and traffic objects all at once. For each task, the design
incorporates a common encoder and three different
In this methodology, a technical framework based decoders. A context tensor is proposed to improve
on the YOLOV4 concept is introduced. This frame- overall performance and computing efficiency by
work focuses on a variety of topics, such as risk facilitating information sharing among activities. DLT-
assessment, object detection, and intent recognition. Net uses the YOLOv3 model for traffic object detec-
Notably, the system uses part affinity fields to add tion, which is a cutting-edge one-stage object detection
human skeletal traits, resulting in enhanced inten- approach. Extensive studies on the BDD dataset show
tion recognition. It also uses LSTM and CNN to that DLT-Net outperforms traditional approaches in
assess vehicle heading, while EfficientNet is utilized these key perception tasks (Qian et al., 2020).
to estimate potentially harmful cars. Furthermore, to
improve risk assessment capabilities, the framework Advantage: its unified design improves efficiency
employs saliency maps generated by the RISE algo- and performance in autonomous driving perception
rithm and explainable AI technology (Guney et al., by recognizing drivable zones, lane lines, and traffic
2022). objects all at the same time.
164 Real-time identification of traffic actors using YOLOv7
Disadvantage: Complex scenarios, such as identifying scenarios. Figure 23.2 represents the architectural
reflected items from traffic signs or dealing with inter- diagram of YOLOv7.
rupted lane lines, may pose difficulties.
Proposed methodology
The study describes a comprehensive autonomous
driving framework that includes four key tasks: Two key elements make up our suggested methodol-
object detection using an optimized YOLOv4 model, ogy and architecture for the real-time detection of
intention recognition based on pedestrian skeleton traffic actors using YOLOv7: Extended efficient layer
features via part affinity fields and CNN analysis, and aggregation networks (EELAN) and a compound
CNN-driven risk assessment for dangerous vehicles scaling technique for concatenation based models.
and traffic light recognition. The YOLOv4 model has
been improved to improve detection accuracy, provid- Extended efficient layer aggregation networks (E-
ing a comprehensive approach to ensuring safe auton- ELAN)
omous driving (Li et al., 2020). Extended efficient layer aggregation networks, or
E-ELAN, are intended to improve network learning
Advantage: It integrates object detection, intention while maintaining the integrity of the initial gradient
identification, and risk assessment to improve auton- path. Expand, shuffle, and merge cardinality tech-
omous driving safety. niques are incorporated into the computing blocks of
Disadvantage: The complexity of the improved PAFs the design to accomplish this. The expand operation
model in the intention recognition component may uses group convolution to expand the channel and
have an effect on computing efficiency. cardinality of the computing blocks. The network is
able to capture a wider variety of features by extend-
ing the channels. Parallel processing and the investi-
Dataset
gation of several feature representations inside each
The enormous collection of images in the traffic computing block are both made possible concurrently
object dataset was specifically picked for the task of by increasing cardinality.
identifying and classifying traffic objects. This data- It makes use of group convolution to keep the orig-
set consists of 4,591 high-quality images that depict inal transition layer of the design. By doing this, it
various real-world traffic situations which have 38 is made sure that the patterns of connectedness and
classes. It is taken from Roboflow where 80% is used information flow between the computing blocks are
for training and the remaining 20% is for testing.
Here, Figure 23.1 represents a sample training image
from the dataset.
Architecture
The YOLOV7 architecture mainly consists of three
parts, i.e., backbone, neck, and head. The backbone
extracts features from the input image, the neck com-
bines features of different resolutions, and the head
generates object detection predictions. This modu-
lar design enables YOLOv7 to efficiently process
input data and accurately detect objects in real-time
Figure 23.1 A sample image from the dataset Figure 23.2 Architecture diagram of YOLOv7
Applied Data Science and Smart Systems 165
maintained. In order to provide seamless information as the backbone network in YOLOv7 offers several
transfer while supporting the enlarged channel and advantages:
cardinality, the transition layer serves as a link between
the earlier computational blocks and succeeding lay- 1. Accuracy: ResNet50 has a strong track record of
ers. It also enables several groups of computational achieving high accuracy on diverse image clas-
blocks to specialize in learning different characteris- sification and object detection datasets, ensuring
tics by utilizing these expand and group convolution reliable results.
procedures. The network can capture and distinguish 2. Efficiency: ResNet50 is known for its relative
traffic actors with increased accuracy because to the computational efficiency, enabling swift training
diversity of feature learning. The shuffling process is and execution, which is crucial for YOLOv7’s
also very important in E-ELAN. According to a pre- real-time object detection design.
determined group parameter, it divides the feature 3. Transfer learning: ResNet50 comes pre-trained
maps produced by the computational blocks into on a vast dataset of images. This pre-training
various groups. By successfully mixing and combin- advantage can be leveraged when training YO-
ing the learned features from many blocks, this shuf- LOv7 on a smaller, custom dataset of images,
fling method promotes feature diversity and guards saving time and resources.
against over-reliance on a single set of computational 4. In summary, ResNet50 is a favorable choice for
blocks. As a result, the model’s ability to generalize YOLOv7’s backbone network due to its ability
and distinguish among traffic actors in real-world cir- to extract rich image features efficiently, leading
cumstances is improved. The E-ELAN process ends to accurate results in object detection tasks.
with the merge cardinality procedure. The merged
feature map with maintained channel numbers is cre- Feature pyramid network (FPN)
ated by joining the shuffled feature maps from several The feature pyramid network (FPN) is a key compo-
groups. nent of the YOLOv7 model. The FPN is responsible
The merge procedure successfully merges the for extracting feature maps at multiple scales, which
many features picked up by several computational allows the model to detect objects of different sizes.
block groups, utilizing their combined knowledge to The YOLOv7 FPN uses top-down architecture with
increase detection precision. EELAN improves the lateral connections. The top-down pathway starts
YOLOv7 architecture overall by allowing ongoing from the highest-resolution feature map and gradually
learning of various features without altering the initial downsizes it while preserving semantic information.
gradient path. Different sets of computational blocks The lateral connections combine the down sampled
can specialize in learning different features thanks to feature maps from the top-down pathway with the
the combination of expand, shuffle, and merge cardi- corresponding feature maps from the backbone net-
nality approaches. work. This results in a set of feature maps at multiple
scales, which are then used by the YOLOv7 head to
ResNet50 predict object bounding boxes and class labels.
ResNet50, CNN architecture, is widely acclaimed The FPN offers distinct advantages over traditional
for its effectiveness in image classification and object single-scale feature extraction methods. Firstly, it
detection tasks. It stands out for its capacity to train enables the model to detect objects of various sizes
deep networks while mitigating the risk of over fit- by providing multiple-scale feature maps. Secondly,
ting. Notably, ResNet50 assumes the role of the back- it enhances object detection accuracy by combin-
bone network in the YOLOv7 model, responsible ing low-level features, rich in spatial information,
for extracting crucial feature maps from the input with high-level features that carry semantic informa-
image. These feature maps serve as the foundation for tion. Thirdly, FPN improves the model’s resilience to
YOLOv7’s head, enabling it to predict object bound- challenges like occlusion and image degradation. In
ing boxes and class labels. The choice of ResNet50 YOLOv7, the FPN is implemented through a series of
as the backbone network for YOLOv7 is strategic. convolutional layers. The initial layer downsizes the
It excels in extracting a diverse and informative set feature map from the backbone network. Subsequent
of features from the input image, a critical factor in layers in this stack handle the task of upsizing feature
the model’s ability to detect and classify objects accu- maps from the prior layers and merging them with
rately. Furthermore, ResNet50 is known for its rela- corresponding maps from the backbone network. The
tive computational efficiency, making it a practical final layer in this stack generates a set of feature maps
choice. This efficiency is particularly important for at various scales, which the YOLOv7 head then uses
YOLOv7, which is designed with real-time object to predict object bounding boxes and class labels.
detection in mind, necessitating a model that can be This approach makes YOLOv7 effective in detect-
trained and executed swiftly. The use of ResNet50 ing objects of different sizes, enhancing accuracy,
166 Real-time identification of traffic actors using YOLOv7
and robustness in the presence of image challenges. RPN classification scores. Bounding boxes are sub-
In conclusion, the feature pyramid network is a criti- sequently produced through RPN bounding box
cal element within the YOLOv7 model, and it greatly regression. The classification layer provides scores to
bolsters the model’s performance across a diverse set the detection layer, which refines the bounding boxes.
of object detection tasks. Both RPN loss and detection loss are included in the
loss module, where the latter combines classification,
Compound scaling method for concatenation-based regression, and object losses, while the former focuses
models on classification and regression losses specific to the
In order to modify the YOLOv7 architecture to meet RPN. These losses, using cross-entropy and smooth
various inference speed requirements, model scaling L1 loss functions, are computed for each image in
is a crucial component. Scaling concatenation-based the batch. This comprehensive module underpins
models, however, presents particular difficulties in YOLOv7’s accurate object detection by enabling
maintaining the ideal structure while attaining the effective region proposals, improved predictions, and
needed scalability. In light of these difficulties, we optimized model training.
provide a compound scaling technique that concur-
rently takes into account the depth and width factors Module-level ensemble (MLE)
of processing blocks and transition layers. It becomes Module-level ensemble (MLE) is a technique employed
especially crucial to preserve the ideal structure when to enhance the performance of object detection mod-
growing concatenation-based models. Performance els by fusing outputs from various modules. In MLE,
shouldn’t be adversely affected by the architecture’s a single module in the model is often replaced with
ability to adapt to variations in depth. In order to multiple parallel modules. These parallel modules
achieve this, our suggested compound scaling strategy generate outputs, which are subsequently fused to
concentrates on maintaining the proportion of input yield the model’s final output. Within the YOLOv7
to output channels while scaling. model, a module-level ensemble layer is integrated
The depth factor describes how many comput- into the neck section of the model. This neck portion
ing units are stacked inside the design. Scaling the plays a role in amalgamating feature maps from both
depth factor alters the in-degree and out-degree of the model’s backbone network and its head.
each layer by increasing or decreasing the number In YOLOv7, the module-level ensemble layer
of computational blocks. The subsequent transition replaces the conventional convolutional layer within
layer’s input-to-output channel ratio is impacted by the neck with an array of parallel convolutional lay-
this modification. To avoid hardware consumption ers. These parallel convolutional layers generate
distortions and guarantee appropriate model param- outputs that are then amalgamated to produce the
eter use, the ratio must be maintained. In addition, final feature maps utilized by the model’s head. This
the width factor, which describes the size of the com- approach enhances the model’s overall performance
putational blocks’ channel, must be changed in pro- in object detection tasks.
portion to variations in depth. The ideal structure of
the original architecture is maintained by scaling the Trainable bag of freebies
width factor, which makes sure that the expanded Trainable bag of freebies (BoF) encompasses tech-
or contracted computational blocks line up with the niques designed to enhance object detection models’
needs of the altered depth. performance without increasing the training cost.
The compound scaling method provides a con- These techniques manifest as trainable modules that
stant ratio between input and output channels all can be seamlessly integrated into existing object detec-
over the architecture by taking into account both the tion models. In the YOLOv7 model, several trainable
depth and width parameters together. This method BoF techniques are incorporated, including:
enables smooth switching between various scaling
factors without impairing the model’s functionality. 1. Cross-module channel communication (C3):
For the model to continue learning and making accu- C3 facilitates inter-module communication by
rate traffic actor distinctions, the ideal structure must sharing channel information, enabling modules
be maintained when scaling. The model’s ability to to learn from each other and enhancing overall
effectively capture and analyze features is maintained model performance.
by the compound scaling strategy, enabling accurate 2. Selective attention module (SAM): SAM enables
and reliable detection of traffic actors in real-time the model to focus on critical parts of the input
circumstances. image, reducing noise processing and conse-
In the YOLOv7 architecture’s detection mod- quently improving model accuracy.
ule, the region proposal network (RPN) generates 3. Efficient channel attention (ECA): ECA em-
anchors based on size and evaluates those using powers the model to discern the significance of
Applied Data Science and Smart Systems 167
In essence, higher average precision (AP) or mean aver- find precision, denoting the proportion of retrieved
age precision (mAP) values indicate superior model instances that are indeed relevant.
performance. On the whole, the results underscore the The graph’s blue line illustrates the precision-recall
YOLOv7 model’s commendable performance on both curve for the model. In contrast, the white line serves
the validation and test datasets, as evidenced by AP as a reference, representing the ideal precision-recall
and mAP scores surpassing 0.5 for all five object cat- curve where precision consistently equals 1. This
egories. Nevertheless, it’s worth noting that there are curve essentially represents perfect performance.
variations in performance across diverse metrics and The mAP@0.5, prominently displayed at the graph’s
object types. For instance, the model exhibits stronger apex, signifies the mean average precision calculated
detection capabilities for cars and people compared to at an IoU threshold of 0.5. It’s a widely used met-
bicycles and motorcycles. ric for assessing the performance of object detection
In Figure 23.4, the x-axis represents recall, which models. A higher mAP@0.5 value is indicative of a
signifies the proportion of all relevant instances suc- more effective model. Examining the precision-recall
cessfully retrieved by the model. On the y-axis, you’ll curve, it becomes evident that the model achieves a
Applied Data Science and Smart Systems 169
high recall while maintaining relatively high preci- Table 23.1 Overview of confusion matrix.
sion. This implies that the model successfully identi-
Predicted True FN FP TN TP
fies a substantial portion of relevant instances without
excessively retrieving irrelevant ones, signifying its Pedestrians Pedestrians 2 10 2000 1990
strong performance.
Vehicles Vehicles 5 5 1995 1990
The mAP@0.5 score of 0.666 is a strong indica-
tion of the model’s ability to perform accurate object Traffic lights Traffic lights 1 1 1998 1998
detection. In Figure 23.5, the model’s performance Stop signs Stop signs 0 0 2000 2000
across each object class is depicted. The matrix Speed signs Speed signs 0 0 2000 2000
rows represent predicted classes, while the columns Buildings Buildings 0 0 2000 2000
denote the true classes. Elements on the diagonal of
the matrix signify the count of correctly classified
objects. For instance, the element at row 0, column 0
represents the number of pedestrians correctly identi- across most object classes. However, it does reveal a
fied as pedestrians. In contrast, off-diagonal elements specific challenge in distinguishing between pedestri-
signify the count of objects incorrectly classified. For ans and vehicles. This difficulty likely arises from the
example, the element at row 0, column 1 indicates the visual similarity between pedestrians and vehicles,
number of pedestrians mistakenly classified as vehi- particularly when observed from a distance.
cles. This matrix provides a comprehensive view of Notably, the model exhibits the highest accuracy for
the model’s performance on individual object classes. object classes like traffic lights, stop signs, speed signs,
The confusion matrix provides an overall posi- and buildings, successfully predicting all instances of
tive assessment of the YOLOv7 model’s performance these categories. In contrast, the model’s accuracy for
170 Real-time identification of traffic actors using YOLOv7
References
Ammar, A., Koubaa, A., Ahmed, M., Saad, A., and Benjdira,
B. (2021). Vehicle detection from aerial images using
deep learning: A comparative study. Electronics, 10(7),
820. [Link]
Bello, I., William, F., Xianzhi, D., Ekin, D. C., Aravind,
S., Tsung-Yi, L., Jonathon, S., and Barret, Z. (2021).
Figure 23.7 (a) & (b) represents the predicted images Revisiting ResNets: Improved training and scaling
generated by the model for the sample input images strategies. Adv. Neural Inform. Proc. Sys. (NeurIPS),
34.
Bochkovskiy, A., Chien-Yao, W., and HongYuan, M. L.
pedestrians and vehicles is comparatively lower. It (2020). YOLOv4: Optimal speed and accuracy of ob-
made incorrect predictions, identifying 10 pedestrians ject detection. arXiv preprint arXiv:2004.10934.
Cao, Y., Thomas, A. G., Jean Yee, H. Y., and Pengyi, Y.
as vehicles and 5 vehicles as pedestrians. These dis-
(2020). Ensemble deep learning in bioinformatics.
crepancies indicate a specific area where the model’s Nat. Mac. Intel., 2(9), 500–508.
performance might benefit from further refinement. Chen, K., Weiyao, L., Jianguo, L., John, S., Ji, W., and Junni,
Bounding boxes are drawn on the input image or Z. (2020). AP loss for accurate one-stage object de-
frame by the algorithm to visually depict the observed tection. IEEE Trans. Pat. Anal. Mac. Intel. (TPAMI),
traffic actors and offer spatial information. These 43(11), 3782–3798.
bounding boxes provide precise information about the Diwan, Tausif, G. Anirudh, and Jitendra V. Tembhurne.
location and size of the discovered items. Figure 23.6 (2023). Object detection using YOLO: Challenges, ar-
illustrates the sample input images provided to the chitectural successors, datasets and applications. mul-
model for prediction. Figure 23.7 shows the sample timedia Tools and Applications. 82(6): 9243-9275.
output images generated by the model based on the He, K., Zhang, X., Ren, S., and Sun, J. (2014). Spatial pyra-
mid pooling in deep convolutional networks for visual
given input images.
recognition. Edited by D. Fleet, T. Pajdla, B. Schiele,
and T. Tuytelaars. (Cham), Lecture Notes in Comput-
Conclusion and future work er Science, 8691. [Link]
10590–12.
In conclusion, the YOLOv7 architecture’s integration Hsu, W.-Y. and Lin, W.-Y. (2021). Ratio-and-scale-aware
of the additional modules (CBS, Mosaic, and ACmix) YOLO for pedestrian detection. IEEE Trans. Im-
and the backbone network (ResNet-50) has shown age Proc., 30, 934–947. [Link]
promising results in the area of automotive vision. TIP.2020.3039574.
The enhanced model collects contextual informa- Li, Y., Hanxiang, W., Minh Dang, L., Tan, N. N., Dongil, H.,
tion, performs multi task detection and classification, Ahyun, L., Insung, J., and Hyeonjoon, M. (2020). A
and extracts features using cross-modal and deep deep learning-based hybrid framework for object de-
CNN. The object classification, object identification, tection and recognition in autonomous driving. IEEE
and image captioning performance analyses have Acc., 8, 194228–194239. [Link]
CESS.2020.3033289.
provided insightful information about the strengths
Lorencık, D. and Zolotova, I. (2018). Object recogni-
and limitations of the model. The performance study tion in traffic monitoring systems. (Kosice, Slova-
revealed competitive item detection accuracy with
Applied Data Science and Smart Systems 171
kia), 277–282. [Link] Yamashita, R., et al. (2018). Convolutional neural net-
8490634. works: An overview and application in radiology.
Mo, Xianglun, Chuanpeng Sun, Chenyu Zhang, Jinpeng Insights Imag., 9, 611–629. [Link] 10.1007/
Tian, and Zhushuai Shao. (2022). Research on Ex- s13244-018-0639-9.
pressway Traffic Event Detection at Night Based on Ye, T., Xi, Z., Yi, Z., and Jie, L. (2021). Railway traffic ob-
Mask-SpyNet. IEEE Access 10 (2022): 6905369062. ject detection using differential feature fusion convo-
Qian, Y., Dolan, J. M., and Yang, M. (2020). DLT-Net: Joint lution neural network. IEEE Trans. Intel. Transport.
detection of drivable areas, lane lines, and traffic ob- Sys., 22(3), 1375–1387. [Link]
jects. IEEE Trans. Intel. Transport. Sys., 21(11), 4670– TITS.2020.2969993.
4679. [Link] 2943777. Ye, T., Zongyang, Z., Shouan, W., Fuqiang, Z., and Xiaozhi,
Reddy, A. Sai Bharadwaj, and D. Sujitha Juliet. (2019). G. (2022). A stable lightweight and adaptive feature
Transfer learning with ResNet-50 for malaria cell- enhanced convolution neural network for efficient
image classification. In 2019 International Conference railway transit object detection. IEEE Trans. Intel.
on Communication and Signal Processing (ICCSP), Transport. Sys., 23(10), 17952–17965. [Link]
0945–0949. IEEE. org/10.1109/TITS.3156267.
Woo, Sanghyun, Jongchan Park, Joon-Young Lee, and In So Zarei, N., Payman, M., and Mohammadreza, S. (2022).
Kweon. (2018). Cbam: Convolutional block attention Fast-Yolo-Rec: Incorporating Yolo-base detection
module. In Proceedings of the European conference on and recurrent-base prediction networks for fast ve-
computer vision (ECCV), 3–19. hicle detection in consecutive images. IEEE Acc., 10,
Xue, Z., Xu, R., Bai, D., and Lin, H. (2023). YOLO-Tea: 120592–120605. [Link] ESS.
A tea disease detection model improved by YO- 2022.3221942.
LOv5. Forests, 14(2), 415. [Link]
f14020415.
24 Revolutionizing cybersecurity: An in-depth analysis of
DNA encryption algorithms in blockchain systems
A. U. Nwosu1, S. B. Goyal2,a, Anand Singh Rajawat3, Baharu Bin Kemat4
and Wan Md Afnan Bin Wan Mahmood5
City University, Petaling Jaya, 46100, Malaysia
1,2,4,5
3
School of Computer Science & Engineering, Sandip University, Nashik, Maharastra, India
Abstract
The rapid advancement of technology has increased the need for robust cybersecurity measures to protect sensitive data
and ensure secure transactions in the digital world. The conventional encryption method has played an appositively role in
the security and privacy of digital systems in the past years. However, emerging cyber threats, such as quantum attacks and
others, pose a looming threat to the security of digital systems. This study explores the innovative approach that leverages
blockchain-based DNA-based encryption algorithms to strengthen the security and privacy of digital systems against these
emerging cyber threats. This paper presents an overview of DNA encryption algorithms and highlights the challenges of
DNA-based encryption algorithms. In addition, this study proposed a blockchain system with DNA-based encryption algo-
rithms to enhance the security and privacy of digital information systems.
Furthermore, the study presented the existing case studies of DNA-based encryption algorithms in different domains of
blockchain systems. Finally, we introduced the challenges of integrating blockchain in DNA encryption. This study concludes
that the proposed solution is more secure and efficient than the conventional DNA encryption approaches, and blockchain
system DNA-based encryption algorithms can potentially revolutionize cybersecurity in emerging digital strategies.
drsbgoyal@[Link]
a
Applied Data Science and Smart Systems 173
e) The comparative analysis shows that the pro- tion in living organisms. Due to its incredible
posed solution is more secure and efficient than density and stability, researchers have explored
the existing system. its potential as a data storage medium. Instead
of traditional electronic storage methods, DNA
The remaining section of this study is organized could store large amounts of information in a
as follows: The background of the work, which tiny physical space.
consists of an overview of DNA encryption and the b) DNA encoding: In DNA encryption, digital data
challenges of DNA-based encryption algorithms. (such as text, images, or files) is converted into
The introduction of blockchain and its operational DNA sequences. This encoding process involves
bases. The literature reviews and related works on mapping binary data (0s and 1s) to DNA bases
types of DNA-based encryption schemes. In addi- (adenine, cytosine, guanine, and thymine). Vari-
tion, it analyses the existing DNA-based algorithms ous coding schemes can be developed to repre-
with its limitations and the analysis of blockchain sent digital data using DNA bases.
systems DNA-based encryption algorithms in differ- c) DNA encryption: Once the data is encoded into
ent domains. The proposed solution – the proposed DNA sequences, encryption techniques can be
algorithm of blockchain-based DNA encryption for applied to enhance security. Traditional crypto-
improved security of digital systems. The challenges graphic algorithms or specialized DNA-based
of integrating blockchain systems with DNA-based encryption methods could be used to protect
encryption algorithms. The case studies of existing the encoded information. The encrypted DNA
DNA-based encryption algorithms leveraging the sequences contain the encoded data in a not di-
blockchain in different sectors and analysis of exist- rectly understandable form.
ing and DNA-based encryption algorithms. Last is the d) DNA decryption: The encrypted DNA sequences
conclusion and recommendation for future research must be decrypted to retrieve the original digital
scope. data; decryption involves reversing the encryp-
tion process, which may require cryptographic
Overview of study keys or specialized DNA-based decryption algo-
rithms. The decrypted DNA sequences are then
DNA-based encryption converted back into binary data. Figure 24.1
DNA encryption is a concept that explores the pos- depicts the cryptographic mechanism of DNA-
sibility of using DNA molecules as a medium for based encryption.
storing and securing digital information (Roy et al.,
2020). It involves converting binary or digital data Challenges of DNA encryption algorithm
into DNA sequences and potentially using DNA- DNA encryption is an emerging field at the intersec-
based encryption and decryption. Here is an overview tion of biotechnology and information security. The
of DNA encryption (Jacob et al., 2013): idea behind DNA encryption is to encode digital
information into DNA molecules, which can then
a) DNA as a data storage medium: DNA is a bio- be stored and processed using biological techniques.
logical molecule that encodes genetic informa- While this concept holds promise for secure data stor-
age, it also presents several significant privacy and
security challenges.
Data integrity: DNA can degrade over time, and envi- The basis of blockchain technology operations is
ronmental factors can impact the stability of DNA- discussed (Swan, 2015; Christidis, 2016).
encoded data. Ensuring the long-term integrity of the
data is a challenge. Decentralization: Traditional centralized systems rely
Data recovery: Developing efficient and accurate on a single authority or intermediary to manage and
methods for retrieving encoded data from DNA mol- validate transactions. In contrast, blockchains oper-
ecules is a significant technical challenge. Data recov- ate on a decentralized network of computers (nodes),
ery processes should be reliable and resistant to errors. where transactions are validated through a consensus
mechanism agreed upon by the network participants.
Biological threats: DNA-based data storage could be
vulnerable to biological attacks, such as introducing Blocks and chains: Transactions are grouped into
harmful biological agents that could compromise the “blocks,” which contain a set of transactions and a
integrity of the DNA data. unique identifier (hash) of the previous block. These
blocks are linked chronologically, forming a “chain”
Scalability: As DNA data storage technologies are still of blocks, hence the name “blockchain.”
in the early stages of development, scalability remains
a concern. Efficient and cost-effective methods for Transparency and immutability: Once a transac-
encoding, storing, and retrieving large volumes of tion is added to a block and that block is added to
data need to be developed. the blockchain, altering or deleting the information
becomes challenging. This immutability is achieved
Cryptography challenges: Developing secure encryp- through cryptographic hashing and consensus
tion algorithms tailored to DNA storage is complex. mechanisms, ensuring that historical records remain
Ensuring that these algorithms are resistant to crypto- tamper-proof.
graphic attacks is essential.
Consensus mechanisms: Consensus mechanisms
Interoperability: Another challenge is ensuring that ensure agreement among participants on the valid-
different DNA data storage systems and platforms ity of transactions (Bamakan et al., 2020). The most
can communicate and exchange data securely. well-known consensus mechanism is proof of work
(PoW), used by bitcoin, which requires miners to
This study will use blockchain technology with
solve complex mathematical puzzles to validate trans-
DNA-based encryption to address the identified
actions. Other mechanisms like proof of stake (PoS),
challenges.
delegated proof of stake (DPoS), and practical byz-
antine fault tolerance (PBFT) offer alternatives with
Blockchain technology
different levels of security and energy efficiency.
Blockchain is a revolutionary technology that has
gained widespread attention for its potential to trans- Security and trust: The decentralized nature of
form various industries and enhance digital trust and blockchain, coupled with cryptographic techniques,
security (Swan et al., 2015; Swan et al., 2017). At its provides a high level of security against fraud and
core, a blockchain is a distributed and decentralized unauthorized access. Transactions are verified by
digital ledger that records transactions across multiple a distributed network, reducing the risk of a single
computers in a transparent, secure, and tamper-resis- point of failure.
tant manner. Figure 24.2 shows the layered diagram Smart contracts: Smart contracts are self-executing
of blockchain technology (Zheng et al., 2018). contracts with the terms of the agreement directly
written into code (Li et al., 2017). These contracts
automatically execute and enforce predefined rules
when certain conditions are met. Smart contracts can
automate various processes, reducing the need for
intermediaries and enhancing efficiency.
Substitute-based scheme: The encoding process in Blockchain system with DNA-based encryption algo-
this technique is carried out using a DNA dictionary rithms
or a look-up table that has been predetermined. In the application of DNA encryption algorithm with
Biological-based scheme: The encryption process is a blockchain system, some work has been done on
carried out using biology-based algorithms. They are this domain on different domains. For example, Kaur
comparatively more secure since they require little et al. (2023) proposed a blockchain-based system
human involvement. for securing and managing healthcare data gener-
Substitute and biological-based scheme: This method ated on cloud networks through DNA cryptography.
performs the encryption using mathematical and bio- Ramaiah et al. (2021) and Arya et al. (2021) designed
logical procedures. The biological operations give a blockchain-based criminal identification using a
an extra layer to the symmetric or asymmetric cryp- DNA encryption algorithm. Table 24.3 analyses the
tographic keys used in mathematical calculations, application of DNA-based encryption algorithms
making them the most secure DNA-based method. with blockchain systems in different domains.
Table 24.1 analyses the types of DNA-based encryp- Based on the limitations of existing literature, this
tion schemes. study will Integrate blockchain-based DNA encryp-
tion algorithms to address the challenges.
Existing DNA-based solutions
Some works have been conducted on the application Proposed solution
of the DNA-based encryption method. For instance
This part presents the proposed blockchain-DNA
Erlich et al. (2017) presented a DNA-based encryption
encryption algorithm to transform cybersecurity.
known as a fountain. This project optimizes digital
data encoding into DNA sequences to enhance data
recovery. It explores efficient DNA-based data storage
techniques, indirectly contributing to encryption and Table 24.2 Analysis of DNA-based encryption solutions
data security. Nandy and Banerjee (2021) presented
a DNA-based image encryption algorithm. The algo- Authors Domain Limitations
rithm aimed to encode images into DNA sequences
Erlich et al., 2017 DNA fountain High latency
and then transmit them securely using DNA’s proper-
ties and proposed a DNA-based data storage system. Nandy et al., DNA-based Lack of
2021 image encryption transparency
This project aimed to store digital data in DNA mol- algorithm
ecules and demonstrated long-term and high-density for secure
data storage potential. Namasudra et al. (2020) pre- transmission
sented a DNA solution. It focused on the encryption Tomek et al., DNA-based data Inadequate
2021 storage security measure
Namasudra et al., DNA-based Long data
Table 24.1 Analysis of different types of DNA-based 2020 encryption in the retrieval time
schemes. cloud computing
environment
Authors Types of DNA- Limitations
based schemes
Jain et al., Substitution- They are highly Table 24.3 Analysis of blockchain system-based on DNA
2014; Hameed based scheme vulnerable to encryption algorithm.
et al., 2018 statistical attacks
Ning, et al., Biological-based It involves higher Authors Domain Limitations
2009; Dhawan scheme computation and is
et al., 2012 time-consuming Kaur et al., 2023 Healthcare Low throughput
Singh et Biological and It involves complex Ramaiah et al., Lack privacy
al., 2017; substitute-based and rigorous 2021
Sukumar et scheme mathematical
al., 2018; calculations Alshamrani et al., IoT Higher latency
Pujari et al., 2021
2018 Liang et al., 2023 Inadequate security
176 Revolutionizing cybersecurity: An in-depth analysis of DNA encryption algorithms
Table 24.4 Case Studies on the integration of blockchain system with DNA based encryption algorithm.
Chernomoretz et al., DNA-based forensic and legal Blockchain was used to store DNA evidence, maintaining
2020 applications using blockchain its integrity and provenance securely. DNA encryption
further protects sensitive genetic information, ensuring only
authorized parties can access the evidence
Kaur et al., 2023 DNA-based secured management Healthcare providers deployed blockchain with DNA
of PHR using blockchain encryption to securely store and share personal health
records. DNA data were encrypted and stored on the
blockchain, ensuring the confidentiality and integrity of
sensitive health information
Ramaiah et al., 2021 DNA-based identity verification DNA samples were used for identity verification on a
using blockchain blockchain. Individuals authenticate themselves by providing
a DNA sample, which is then encrypted and stored on the
blockchain, enhancing security for digital identities
Chernomoretz et al., DNA-based genomic data The blockchain’s decentralized and immutable nature enables
2020 privacy and ownership using individuals to retain ownership and control over their
blockchain genomic data. The encrypted DNA sequences were stored on
the blockchain, and individuals could grant specific access
permissions to researchers, doctors, or institutions. Smart
contracts facilitated data sharing while ensuring privacy and
allowing data owners to revoke access anytime
Liang et al., 2023 DNA-based pharmaceutical Blockchain establishes an auditable and tamper-proof record
research and intellectual property of research milestones: the DNA sequences and intellectual
using blockchain property. The DNA encryption algorithms safeguard
proprietary genetic information while allowing secure
collaboration between different parties. Smart contracts
automate royalty distribution and licensing agreements,
reducing disputes and enhancing stakeholder trust.
178 Revolutionizing cybersecurity: An in-depth analysis of DNA encryption algorithms
2
Faculty of Information Technology, City University, Petaling Jaya, Malaysia
3
School of Computer Sciences and Engineering, Sandip University, Nashik, India
4
Ramrao Adik Institute of Te1chnology, Navi Mumbai, India
Abstract
In this research paper, the authors have focused on predicting indicators of recession conditions using public opinion-based
platforms and newspaper resources. A three-staged data science pipeline is created which involves data collection from vari-
ous platforms, data filtering, and data cleaning process. In the last stage of the pipeline, we analyzed the data and generated
insights from it. In the data collection process, we have collected real-time based data from public-opinionated social media
platforms like Twitter and Reddit. Additionally, New York Times articles have been collected for the purpose of a newspaper-
based platform We have performed natural language processing (NLP) methods like keyword analysis, word-frequency
analysis, and sentiment analysis to compare the change in the attributes of data over time. The results suggest that NLP
techniques tools can be used to prove the short- and long-term indicators of recession conditions and inflation reasons across
the globe on public-opinionated platforms and newspaper articles.
Keywords: Inflation prediction, natural language processing, New York Times, recession, Reddit, Twitter
1
drsbgoyal@[Link]
drsbgoyal@[Link]
a
Applied Data Science and Smart Systems 181
This study, in the context of the convergence of York Times newspaper articles are used to evaluate
machine learning (ML) and natural language process- the post-covid economic recession conditions across
ing (NLP) techniques, seeks to uncover potential pat- the globe. Barhanpurkar et al.’s study (2023) in
terns and relationships between online discourse and which the authors have performed sentiment analy-
economic trends. The entire paper is divided into 5 sis, entity recognition, and topic modeling. The sen-
sections and their content is as follows: Introduction, timent analysis and NLP processing techniques are
Related Work, Proposed Methodology, Results and used to gain insights from the corona Covid-19 pan-
Discussion and Conclusion and Future Work. demic outbreak on NY Times articles (Tunca et al.,
2023). In Table 25.1, the different studies show the
Related work use of Twitter, Reddit, and the New York Times which
shows the broad spectrum of domains in which the
Twitter is one of the important data sources for data is used.
analyzing several factors during corona Covid-19
pandemic situation since October 2019. A detailed
Proposed methodology
analysis of the Covid-19 vaccine been carried out
based on 4 million tweets and the parameter were In Figure 25.1, the research methodology employed
discovered as the number of tweets who are against during the research is described. In the initial step of
the vaccine (anti-vaccine) and who support the vac- data collection and storage, data is obtained from
cine (pro-vaccine) (e) (Yousefinaghani Samira et al.,
2021). Similarly, the key indicators of economic ten-
sions and war in Ukraine are analyzed using Twitter
Table 25.1 Comparative analysis of different studies
based on 42 million tweets. It also highlights the associated with Twitter, Reddit, and the New York Times.
impact of war on the US dollar value and crude oil
values across the globe (Polyzos, 2022). Twitter data Study Year Platform Domain
quality standards practices are one of the crucial fac-
tors which handle the further analytical process and Edo-Osagie 2020 Twitter Public Health
et al. care
correct results (Salvatore et al., 2021). Feng Yunhe
et al. (2023) have supplied the impact of chat-GPT Malik et al. 2019 Twitter Education
on streaming media using social media platforms Baker et al. 2021 Twitter Economics
such as Twitter and Reddit. The study has been col- Pirina et al. 2018 Reddit Public
lected on real-time analysis where the response time. healthcare
The Reddit forum data is used to gather the stu- Karpenko 2021 Reddit Personal
dents’ requirements for changes in the infrastructure et al. banking
requirement during corona Covid-19 pandemic out- Ireland et al. 2023 Reddit Employment
break (Feng et al., 2023). Additionally, the student’s sector
mental health parameters are also evaluated on the Alieva 2023 New York Times Government
loan debts using Twitter and Reddit platforms (Sinha Rodden 2021 New York Times Economics
et al., 2023). The Reddit posts are majorly used for Costola et al. 2023 New York Times Economics
topic modeling because the users post comments in
Ujewe 2023 New York Times Humanities
the sub-reddit group (Bonifazi et al., 2023). The New
three various sources, including tweets from Twitter collected, and the y-axis represents the number of
API, sub-reddit comments from Reddit API, and New tweets we have collected on dates. We have collected
York Times articles and headlines from New York 155,218 tweets, 240,079 sub-reddit articles (Figure
Times Archive API. The streaming data from the 25.2(b)) and 616,895 comments (Figure 25.2(c)) col-
sources is being stored in the MySQL database using lected using Twitter, New York Times, and Reddit,
the MySQL Connector. After the completion of the respectively.
data collection and storage phase, the data cleaning For recession tracking, the use of real-time data is
step is carried out to get the data ready for NLP oper- a fundamental step in the evolution of social media
ations. Punctuation removal, tokenization, and stem- platforms such as Twitter, Reddit, and The New
ming are some of the techniques used to refine the York Times. These benefits include quicker response
data. In the data visualization and analytics stage, the to economic events, early identification of sentiment
processed data is finally used to generate insights on shifts, capture unconventional indicators, interactive
the recession condition, economic crisis, and inflation analysis with the public, complementary insights to
markers. The methods used for generating insights traditional indicators, and fostering innovative data
consist of time series analysis, memory usage analy- analytics techniques. It shows the data collection steps
sis, word cloud analysis, word frequency analysis, and that are carried out from the different data sources.
lastly, the sentiment analysis. Time series analysis is We have used open-sourced platforms Twitter API,
used to identify trends in the data and cyclical pat- Reddit API and NYTimes API for data collection
terns that focus on how the nature of the data has (Table 25.2).
changed over time. Furthermore, memory usage anal- Figure 25.3 consists of a graph in which the x-axis
ysis optimizes data processing workflows by ensuring represents the number of days we collect the data
the scalability of computational resources. The word
cloud analysis simultaneously visualizes the major
themes in the data and growing trends are revealed by
word frequency analysis. Thus, word cloud analysis
and word frequency analysis are the two major steps
of keyword analysis for large scale textual datasets.
from the data sources. The y-axis represents the between 10 and 30. The word frequency analysis is
dataset size which is the size of data collected for the the first step of keyword analysis. It provides insights
corresponding days. This graph provides an insight about the number of keywords that can be extracted
into the memory usage done by each dataset as data from the entire dataset.
sources like Reddit, Twitter, and NYTimes produce Figure 25.5(a) consists of a word-cloud for the
hugely various kinds of data. data collected from the Twitter data. The words that
Figures 25.4 (a–c) consist of a graph representation are commonly used represent the words related to
in a tweet (Twitter), Reddit comment (Reddit API), the recession. Words like “government,” “inflation,”
and NYTimes article abstract, respectively where the
x-axis represents the number of words in each record
and the y-axis represents the number of occurrences.
In the Twitter data, maximum occurrences of words
can be obtained in the range of 20–40. The Reddit
data contains the maximum number of occurrences
in the range of 5–15.
Additionally, the NYTimes data collected in the
abstract form of the article also contains a range
Conclusion
In result and discussion section, we have determined
the major insights which we have obtained from the
data during data collection, data pre-processing, data
filtration and data visualization steps. Because of these
results, we have drawn some conclusions for the pro-
posed research question (RQ) this has been mentioned
in the introduction section. In response to research
question 1 (RQ1), we have performed the keyword
analysis based on the word cloud and character fre-
quency analysis. We have found that public opinion-
ated datasets (Twitter and Reddit) keywords changed
more frequently as compared to the NYTimes articles
data. More political and geographical significance has
Figure 25.6(b) Sentiment score of the Reddit posts been observed in the information in the news articles
dataset as compared to the Twitter and Reddit data.
Additionally, the frequency of the characters per sum-
marized news article is more as compared to public
opinionated datasets.
For the RQ2, the authors analyzed this research
question based on the sentimental analysis score cal-
culated for each tweet, subreddit comment, and news
article summary. We have used the sentimental analy-
sis correlation scale for the analysis. The sentimental
score ranges from -1 to +1 where -1,0 and +1 sig-
nify negative, neutral, and positive, respectively. The
New York Times article data show extreme negativ-
ity and positivity compared to the public opinionated
dataset. In the Twitter dataset, the sentimental score
shows normal distribution which concludes that the
user provides several types of opinions on the reces-
Figure 25.6(c) Sentiment score of the Reddit posts
sion topic. In the Reddit data, we have observed
that Reddit users follow the trend of an extremely
negative score and positive comments are uniformly
Figure 25.6(a) consists of a graph in which the distributed on the recession topic. The research ques-
x-axis represents the sentiment analysis score ranges tion 3 (RQ3) particularly focused on the influence of
of different tweets related to the recession. The score the social networking sites (Twitter and Reddit) on
range 0.0–0.25 has the highest occurrence in this topic or entity. The filtered data result shows that the
graph. Figure 25.6(b) consists of a graph in which user response rate and user-engagement increased on
the x-axis represents the sentiment analysis score recession topic.
186 Exploring recession indicators: Analyzing social network platforms and newspapers
Abstract
This study explores the impact of outcome-based accreditation and assessment, encompassing the National Board of Ac-
creditation (NBA), National Assessment and Accreditation Council (NAAC), and National Institutional Ranking Frame-
work (NIRF), on engineering education. These bodies aid students in achieving excellence in higher education standards,
with each entity utilizing distinct criteria for evaluating engineering programs’ credibility. The primary aim of this research
is to assess the effectiveness of these ranking methods in enhancing education quality, particularly in aiding private engineer-
ing institutions to improve their reputation. The study tests the null hypothesis, assuming no effect, against the alternative
hypothesis. Evaluation metrics include teaching, learning and resources (TLR) scores, research and professional practice
(RPC) scores, graduation outcome (GO) scores, outreach and inclusivity (OI) scores, and perception score for individual
colleges. The model’s efficiency is determined using the standard strategic indicator: root mean square error. A low value of
this indicator implies efficient NIRF rank prediction by the model. The p-value for the research project was determined to be
0.0466922, whereas the p(x) F-value was 0.953308. Therefore, a revised explanation that considers this is appropriate, such
as the speculation that the NIRF rating will significantly impact schooling.
a
drsbgoyal@[Link]
188 NIRF rankings’ effects on private engineering colleges for improving India’s educational system
are some instances of assumptions: the selection of equal opportunities, leading to positive perceptions
samples has nothing to do with chance. The popula- of institutions implementing them. Outreach and
tion from which the sample was drawn is assumed to inclusivity can be improved by promoting diversity,
have a normal distribution, and each standard devi- equity, and inclusion on campus and engaging with
ation should be equal (1 = 2 =... = k). Two further the local community. Computational approaches can
assumptions have been made about what is occurring. be used to analyze the NIRF rankings and their effects
When there is a significant difference between the two on private engineering colleges in India. For example,
groups being compared, the assumption plays a more data mining techniques can be used to identify pat-
significant part in the research. terns and trends in the rankings over time. First, the
The outcome-based education (OBE) framework study obtains test data from the open-source Kaggle
emphasizes student outcomes, promoting research, to validate and analyze the model’s accuracy. Then,
and improving graduation outcomes while fostering the study analyzed the test data using the instruc-
inclusivity. Continuous quality improvement (CQI) tions provided by the same open source from which
focuses on ongoing enhancements in teaching and we obtained the test data. Table 26.1 depicts the data
learning, ensuring quality education accessible to all, set, while Figure 26.1 illustrates how the research was
which positively influences NIRF rankings. Six Sigma’s conducted.
quality control and TQM’s continuous improve-
ment efforts enhance teaching, research, graduation Results
outcomes, and inclusivity, favorably impacting the
rankings. In summary, these frameworks contribute This study aims to determine how effective the numer-
to elevating NIRF rankings by emphasizing quality, ous education rules and accreditation systems, such
research, and outcomes while ensuring inclusivity and as NIRF, are at enhancing the overall performance
Aspect Description
Outcome-based education Emphasizes on the importance of student outcomes through the curriculum, teaching
(OBE) methods, and assessments, thus improving the teaching, learning, and resources.
Promotes research as a part of the curriculum, encouraging students to apply knowledge
professionally. Main focus is to produce graduates who can meet specific outcomes,
leading to improved graduation outcomes. Encourages diversity and equal opportunities
for all students, improving outreach and inclusivity
Continuous quality Focuses on continuous improvement in teaching and learning resources by regularly
improvement (CQI) assessing and updating them framework
Promotes research by focusing on continuous improvement and innovation in
professional practice
Continuous improvement approach ensures better graduation outcomes
Ensures that quality education is accessible to all, thus improving outreach and
inclusivity
The perception of an institution implementing CQI is generally positive due to its focus
on quality improvement
Six sigma framework Aims to improve teaching and learning by reducing defects and variability in educational
processes
Encourages research by promoting the use of data-driven methodologies in professional
practice
Total quality management TQM focuses on improving the quality of teaching and learning resources through
(TQM) framework continuous feedback and refinement
of private engineering educational institutions. A 2.2983 falls outside the 95% confidence interval [-:
one-way analysis of variance was performed utiliz- 2.2611], which we already know. The force of the
ing the F distribution and a df value of 5,192. This hit and the magnitude of the observed effect, both
allowed for comparing how the NIRF ranked all 33 denoted by f, are regarded as the average value (0.24).
states from 2016 to 2021 (being on the correct path). This indicates that the difference between the means
Before examining the test findings, it is assumed that is comparable to a moderate difference. The variable’s
both the null hypothesis H0, stating that the ranking value was determined to be 0.056. Consequently, the
system has no effect, and the alternative hypothesis group accounts for 5.6% of the total standard devia-
H1, stating that the NIRF schooling system has a sig- tion (similar to R in the linear regression). Upon
nificant effect, are true. The ranking system has no learning about the Turkey HSD and the Turkey
influence, as the null hypothesis H0 states. According KRAMER, it becomes apparent that this is the case.
to the alternative hypothesis H1, the NIRF education When comparing the means of two groups, there is no
system has a significant impact. discernible difference between them. The average of
Examining Table 26.2 reveals that the results sup- multiple groups can be significantly different from the
port the validity of the H hypothesis. Hypothesis H average of a single group or from any other collection
cannot be valid given the low p-value. It has been of means. This is not the same as claiming that the
brought to our attention that the averages of some of average of any set of means cannot have a significant
the groups are significantly different from those of the value difference. With a test power of 0.7696 and a
others. In other words, there is a substantial differ- medium a priori power, it is possible to demonstrate
ence between the means of some categories, and this that the null hypothesis H0 is false during the vali-
difference could be considered statistically significant dation phase. We can compare the test power to the
due to its magnitude. In Table 26.2, the study can see priori power to accomplish this goal.
that the p-value for the study project is 0.0466922 In contrast, the investigation results led the
and that [p(x F)] is 0.953308. In other words, the researchers to conclude that variances are equal when
likelihood of committing a type 1 error, which in variance equality is considered. The tool utilized
this instance would be incorrectly excluding an H, is Levene’s test to determine whether the differences
relatively low: 0.04669 (4.67%) When the p-value is were comparable. We are operating on the assump-
smaller, it indicates that there is more evidence that tion that disparities in population means are, for the
H exists. In addition, the evidence implies that F = most part, comparable. The value of p is 0.103. When
192 NIRF rankings’ effects on private engineering colleges for improving India’s educational system
it does not assume that all groups have the same level
of variation. This is true when the group sizes being
compared are identical (the difference between the
larger and smaller groups is 1). It seems plausible to
conclude that the study’s conclusions are accurate. In
the framework of the presumption of normality, the
Shapiro-Wilk test was utilized to support the premise
of normality (α=0.05). It is anticipated that at least
30 individuals would comprise each category’s sam-
Figure 26.2 Confidence intervals ple. Each sort of analysis has a distinct visualization
technique. In addition to the more typical confidence
intervals, the F distribution curve, the histogram, and
the power F distribution are all examined. The fol-
lowing section depicts a Figure 26.2 depicts the con-
fidence intervals, whereas Figure 26.3 depicts the F
distribution curve, histogram in Figure 26.4 and the
power F distribution curve shown in Figure 26.5.
Conclusion
The NIRF rankings have been shown to affect India’s
private engineering colleges significantly. The authors
Figure 26.3 Distribution curve of this work use different computer methods to learn
more about this. The Indian Ministry of Human
Resource Development established the NIRF rank-
ings. In particular, tests derived from statistical varia-
tion analysis are utilized in the inspection process.
NIRF is being thought about because outcome-based
accreditation and assessment have helped engineering
education professionals in the past (Khatoon et al.,
2022). In addition, they assist students in meeting the
Excellence in Higher Education program’s standards.
Several criteria, such as those utilized by the NBA, the
NAAC, and the NIRF, determine the most trustworthy
engineering programs. The main goal of this study is
to find out if and how this ranking method could help
Figure 26.4 Histogram improve the education system as a whole. The NIRF
was developed to enhance the standing of privately
funded institutions that offer engineering degree pro-
grams. The project’s objective was to enhance these
institutions as a whole. This was the most significant
objective of the project. The “null hypothesis,” which
states that there is no impact, is given the benefit
of the doubt during scientific inquiry. On the other
hand, it is demonstrated that the alternative hypoth-
esis that there is an effect is wrong. Due to previous
findings, this conclusion may be plausible. We now
know that the p-value for this study is 0.0466922,
the F statistic is 2.298, and the p(x) F value for this
Figure 26.5 Power F distribution curve study is 0.953308. These are the values that the study
established. The study can also view images depicting
the outcomes of this inquiry. Users can access various
utilizing Levene’s test, it is prudent to presume that graphical tools, such as confidence intervals, distribu-
the force is modest (0.77). It is simple to compare the tion curves, histograms, and the power F distribution
categories because their sizes are comparable. Since curve. All of these tools are designed to help them
the ANOVA test employs a distinct statistical model, interpret the data.
Applied Data Science and Smart Systems 193
Additionally, there are other visual tools avail- Nancy, W., Parimala, A., and Merlin Livingston, L. M.
able for use (Dutta et al., 2022). For example, one (2020). Advanced teaching pedagogy as innovative
new explanation that considers this is the idea that approach in modern education system. Proc. Comp.
the NIRF rating will have a big effect on how likely Sci., 172, 382–388.
Surekha, T. P. and Shobha, S. (2020). Enhancing the quality
someone will get an education. This is an example of
of engineering learning through skill development for
the type of acceptable explanation.
feasible progress. Proc. Comp. Sci., 172, 128–133.
Khattar, N., Singh, J., and Sidhu, J. (2020). An energy effi-
References cient and adaptive threshold VM consolidation frame-
work for cloud environment. Wire. Per. Comm., 113,
Nassa, Anil Kumar, Jagdish Arora, Priyanka Singh, J. P. 349–367.
Joorel, Kruti Trivedi, Hiteshkumar Solanki, and Ab- Møller-Skau, M. and Fride, L. (2022). Arts-based teaching
hishek Kumar. (2021). Five Years of India Rankings and learning in teacher education: “Crystallising” stu-
(NIRF) and its Impact on Performance Parameters dent teachers’ learning outcomes through a systematic
of Engineering Institutions in India. Pt. 2. Research literature review. Teach. Teach. Educ. 109, 103545.
and Professional Practices. DESIDOC Journal of Li- Bedi, P., Pushkar, G., Shivani, D., and Neha, G. (2020).
brary & Information Technology, 41(2), 116–129. Smart contract based central sector scheme of scholar-
DOI:10.14429/DJLIT.41.02.16674 ship for college and university students. Proc. Comp.
Vasudevan, N. and T. Sudalaimuthu. (2020). Development Sci., 171, 790–799.
of a common framework for outcome based accredita- Madheswari, S. P. and Uma Mageswari, S. D. (2020).
tion and rankings. Proc. Comp. Sci., 172, 270–276. Changing paradigms of engineering education - An
Gupta, S. (2019). Chan-vese segmentation of SEM ferrite- Indian perspective. Proc. Comp. Sci., 172, 215–224.
pearlite microstructure and prediction of grain bound- Reddy, K. S., En, X., and Qingqing, T. (2016). Higher educa-
ary. Int. J. Innov. Technol. Explor. Engg., 8(10), 1495– tion, high-impact research, and world university rank-
1498. ings: A case of India and comparison with China. Pac.
Sengupta, I., Chandan, K., Niloy Kumar, B., and Subir, G. Sci. Rev. B Hum. Soc. Sci., 2(1), 1–21.
(2022). Automated student merit prediction using Kumar, V. and Preedip Balaji, B. (2021). Correlates of the
machine learning. 2022 IEEE World Conf Appl Intel. national ranking of higher education institutions and
Comput. (AIC), 556–560. funding of academic libraries: An empirical analysis. J.
Mondal, B., Debkanta, C., Niloy Kumar, B., Pritam, M., Acad. Librarian., 47(1), 102264.
Sanchari, N., and Subir, G. (2022). Review for meta- Johnes, G., Jill, J., and Swati, V. (2022). Performance and
heuristic optimization propels machine learning com- efficiency in Indian universities. Socio-Econ. Plan. Sci.,
putations execution on spam comment area under 81, 100834.
digital security Aegis region. Integ. Meta-Heur. Mac. Khatoon, Fahmida, Manish Kumar, Ayesha Akbar Khalid,
Learn. Real-World Optim. Prob., 343–361. Amal Daher Alshammari, Farida Khan, Rashid D.
Anoop, C. A. and Pawan, K. (2013). Application of Taguchi Alshammari, Zahid Balouch et al. (2022). Quality
methods and ANOVA in GTAW process parameters of life during the pandemic: a cross sectional study
optimization for aluminium alloy 7039. Int. J. Engg. about attitude, individual perspective and behavior
Innov. Technol. (IJEIT), 2(11), 54–58. change affecting general population in daily life. 6th
Bhat, S., Sathyendra, B., Ragesh, R., Rio D’Souza, and Binu, Smart Cities Symposium (SCS 2022), 379–383.
K. G. (2020). Collaborative learning for outcome Dutta, P.K., Bose, M., Sinha, A., Bhardwaj, R., Ray, S., Roy,
based engineering education: A lean thinking ap- S. and Prakash, K.B. (2022). Challenges in metaverse
proach. Proc. Comp. Sci., 172, 927–936. in problem-based learning as a game-changing virtu-
Jadhav, M. R., Anandrao, B. K., Satyawan, R. J., and Ma- al-physical environment for personalized content de-
hadev, S. P. (2020). Impact assessment of outcome velopment 6th Smart Cities Symposium (SCS 2022),
based approach in engineering education in India. 417–421. doi: 10.1049/icp.2023.0641
Proc. Comp. Sci., 172, 791–796.
27 Analysis of soil moisture using Raspberry Pi based on IoT
Basetty Mallikarjuna1, Sandeep Bhatia2,a, Amit Kumar Goel3,
Devraj Gautam4, Bharat Bhushan Naib5 and Surender Kumar6
1
Department of Information Technology, Institute of Aeronautical Engineering, Dundigal, 500043
2,5
School of Computing Science and Engineering, Galgotias University Greater Noida, Uttar Pradesh, India
3
School of Engineering and Technology, Apeejay Stya University Sohna Gurugram, India
Department of Electronics and Communication Engineering, Dr. Akhilesh Das Gupta Institute of Technology and
4,6
Abstract
Agriculture accounts for a large portion of India’s economy However, farmers often lack access to essential farming equip-
ment. They are confronted with issues, for example, a lack of soil fertility or insufficient soil water treatment hydration and
others. The internet of things (IoT) is an interconnection of devices which have unique identity and able to share data in real
time. An autonomous farming system can be constructed using IoT to reduce water waste and boost crop yields. Using a soil
moisture sensor with a Raspberry Pi, Pico module and node MCU, the water level is monitored and recorded. The soil retains
information about its moisture levels over time. The node MCU takes data in analog values and analyze it before sending it
to the Raspberry Pi Pico via the telegram program, where the user can observe the current moisture level. This is crucial as
plants require adequate water. For a decent yield, it needs to be watered at a precise time. The developed system able to send
information to a remote location related to soil in real time by using Raspberry Pi module. Optimization of the information
related to soil received from Raspberry Pi will be carried out to enhance productivity.
Keywords: Internet of things, soil moisture sensor, Raspberry Pi Pico, node MCU, agriculture
sandeepbhatia1711@[Link]
a
Applied Data Science and Smart Systems 195
expenditure with more feasibility. Contrarily, it may information unit (WIU) and wireless sensor unit
have the ability to provide technological momentum (WSU) (Holliday et al., 1990). Zigbee-based wire-
for the revival of the world economy (Mallikarjuna, less sensor networks (WSN) are used to accom-
2020). IoT innovation is now being used in a variety plish this. There is a standard at Western Illinois
of industries, it also addresses several challenging sub- University (WIU) for providing data to online
jects (Mallikarjuna, 2022). IoT and agriculture will services. Even though this technology succeeds in
be a winning combination and definitely contribute automating tasks, there are drawbacks (Knight et
to the resolution of present horticulture framework al., 2017). The solenoid valve in this study regu-
inefficiency issues, as well as the rapid and efficient lates the opening and closing of the valve, while the
enhancement of farming. The water level is monitored microcontroller manages the signal in the sensor-
using the Raspberry Pi Pico module and node MCU in based IoT irrigation system. The water process is
conjunction with a soil moisture sensor (Mallikarjuna, initiated and the water flow is regulated in response
2020). Agrarian data may also be obtained through to changes in the ambient temperature and humid-
the crops, related resources, and equipment being sci- ity. When the humidity is low, water your plants;
entifically monitored using IoT and Big Data process- when the moisture content is back to normal, stop
ing technology to optimize better crop management watering. Microcontrollers and GSM are interfaced
(Mallikarjuna, 2022). via MAX232 (Knight et al., 2015). Although auto-
There is a monitoring of the water level and matic water use is the desired outcome, this tech-
Raspberry Pi Pico module and node MCU with a soil nology accomplishes it in a more sophisticated
moisture sensor were used to record this video. The manner (Marthaler et al., 2023). In order to opti-
soil carries information on the moisture state of the mize agricultural water use, this study presents an
soil over some time (Mallikarjuna, 2022). The node automatic water management and field monitoring
MCU takes data in analog values and analyses it system (Knight et al., 1992). The automatic systems
before sending it to the Raspberry Pi Pico via the tele- are based on WSN (Brar et al., 2022). A wireless
gram program, where the user can observe the current network of temperature and humidity sensors that
moisture level (Srinivasan et al., 2022). This is signifi- are placed in the field is a feature of the system.
cant since the plant needs water at a precise period to Data is transferred from the sensor to the microcon-
produce optimum yield outcomes. Manually measure troller (Khattar et al., 2020) via the Zigbee protocol
soil dampness since mistreating this tensiometer might (Attema, 2007). Farmers receive information via a
take a long time (Sandeep et al., 2023). We would PIC 18F77 microcontroller with GPRS support and
therefore need a system that can track the amount of a GSM modem (Singh et al., 2020). The PIC micro-
water (moisture) in the soil in a very short amount of controller’s usage of RISC determines how long the
time while also being simple to operate (Supriya et al., program will run (Gutiérrez, 2013).
2022). increased yields; increased accuracy by reduc- The author Teka (2019), has developed an intelli-
ing “skipping” (omissions) and “doubling” (repeated gent drip irrigation system that is controlled by an
applications – overlaps) between adjacent rows in the ARM9 processor and is automatic. The soil’s pH and
field (Sathish et al., 2021). nitrogen content are continuously shifting through-
out this process. GSM modules are employed for situ-
• Enhanced productivity: faster working speeds are ational monitoring and control (Anand et al., 2015,
feasible; 2017). In this instance, the soil’s moisture content
• Increased safety; and the capacity to work at is estimated using an acoustic-based technique. This
night and in low-light conditions. strategy’s primary goal is to speed up the measure-
ment of soil moisture (Prasad et al., 2018). According
The current publication builds on the work of the to Li et al. (2016) and 2020, the two primary vari-
aforementioned authors by attempting to verify and ables in this process are soil water saturation and
quantify the predicted economic savings. sound speed. It was thus discovered that the speed
The following is how this research article is struc- of sound varies with soil type and that it diminishes
tured: A brief overview of prior studies is presented; with increasing soil moisture. Farmers are unable
the proposed approach is described in depth; the to use this technique, despite the fact that it can be
experiment is implemented, and the results are dis- used to quickly determine the moisture content of a
cussed; and finally the proposed work is concluded. soil sample (Mallikarjuna et al., 2022). The system
was developed using an automated irrigation system
Literature review equipped with a soil moisture sensor and a smart-
phone or tablet (Mallikarjuna, 2022). With the help
This unit transmits temperature and humidity of this technology, people can conserve water and
data via a radio transceiver-equipped wireless increase water duration control. For calibration, the
196 Analysis of soil moisture using Raspberry Pi based on IoT
model was tested with a range of crops and soil sam- this paper comes from a questionnaire survey as well
ples at varying soil moisture levels. as a FADN agricultural product. Based on their busi-
Nevertheless, by utilizing more soil samples from ness structure, the businesses studied can be classified
various locations and climates, this outcome could be as natural persons or legal entities. Natural person
enhanced. We’ll look at other soils in addition to soil businesses made up 14.3% of the agricultural enti-
moisture. Singer et al. (2021) investigates electronic ties studied. Legal entities accounted for 85.7% of the
gadgets that gather and transmit physical data to users total.
via IoT. Finding quick fixes for issues and presenting System starts with deployment of sensor circuitry
workable solutions are the goals of this project. IoT- consist specific set of sensors and Raspberry Pi Pico
based smart agriculture can make use of 4G to 8G collect data from farming land, if data gathered suc-
communication (Sandeep et al., 2023). According to cessfully then it is transmitted to remote location.
Kai et al. (2022), security-enhanced features are cru- After that there is optimization of the crop data to
cial for safeguarding sensitive data in IoT-based smart increase crop yield and finally important information
farming. It is possible to use heterogeneous nodes to regarding published data published on telegram as
gather data from agricultural fields (Shabana et al., shown in Figure 27.1.
2023). The plant part is the first section, and it has sensors
for detecting the environment around the plant. The
Objectives DHT11 soil sensor is a cheap soil moisture sensor that
can be used to track soil moisture, just like other soil
This paper is aimed to design a framework for the moisture sensors. The moisture sensor outputs a high
analysis of soil moisture using Raspberry Pi-based level when the soil is dry, and a low level otherwise.
on IoT. The objective of paper is to construct a sys- In contrast to other soil moisture sensors, this one has
tem which can be utilized to optimize soil parame- a variable sensitivity and gives the IoT hub raw data.
ters, reduce water waste and boost better crop yield.
Another objective is to published data on telegram IoT hub
which enable farmers to increase crop yield. You can connect your devices to the internet via IoT
hub service which has devices to communicate in both
Proposed methodology directions. The IoT hub connects various services to
the real world. The data from the sensors is updated
Numerous technologies are at one’s disposal, such on a regular basis. The primary function of the IoT
as video smart working, code management systems, hub is to monitor and connect all connected IoT
smart home offices, electronic products, and more. devices.
External assistance has a lot of advantages. Such as
being RFID (Radio Frequency Identification) labeled, Analytics in streams
being shrewd, and so on. Background information for The IoT hub offers the service of stream analytics,
which is necessary for data transmission from the IoT
hub. Basically, this service’s fundamental character-
istic is its capacity to stream millions of records per
second, or millions of pieces of data, in real time.
Figure 27.2 shows the system execution in which
system starts collecting data and transmit it to remote
location for optimization and then published on tele-
gram. Also, Figure 27.2 shows the system descrip-
tion, the Raspberry Pi Pico, node MCU, and soil
moisture sensor are among the components used in
this system.
Figure 27.3 Circuit details of Raspberry Pi Figure 27.5 Node MCU (ESP 8266)
198 Analysis of soil moisture using Raspberry Pi based on IoT
stores information about the soil’s moisture status Figure 27.7 depicts the yield and response graph
throughout time. The Node MCU takes data in ana- displaying water and yield information. A signal is
logue values and analyses it before sending it to the sent to the Raspberry Pi Pico board from the out-
Raspberry Pi Pico via the telegram program, where put. The Raspberry Pi is the system’s brain. The
the user can observe the current moisture level. This Raspberry Pi now boasts a plethora of updated fea-
is significant since the plant requires a lot of water. tures. Utilizing the sensor attached to the Raspberry
For a decent yield, it needs to be watered at a precise Pi board, determine the resistance difference. The
time. humidity is determined by the cold signal circuit
with potentiometer if the comparator’s output is
Test strategy high. The Raspberry Pi board receives the output
All objects are connected to the IoT network in order signal. The Raspberry Pi is the system’s brain. The
to achieve interconnectivity and data transmission Raspberry Pi now boasts a plethora of updated
capabilities over the internet and traditional media. features. The Raspberry Pi board receives the out-
Following that, it was perceived as a wide range of put signal. The Raspberry Pi serves as the system’s
high-end devices and workplaces that were unaf- brain.
fected by the district. “External enablement” and
“internal intelligence” were included. Examples of
devices used to cultivate inner wisdom include digi-
tal control systems, smart home offices, smart video
roles, mobile devices, technology systems, meters, etc.
Outside enablement refers to a wide range of benefits,
such as items branded with RFID (Radio Frequency
Identification), as well as smart people and cars
equipped with wireless terminals.
watering system was developed without the usage Surya Prasad, P. and Prabhakara Rao, B. (2016). Curvelet
of a Raspberry Pi and instead made use of a wire- transform based statistical pattern recognition sys-
less sensor network and GPRS module. Surya et al. tem for condition monitoring of power distribution
(2018) proposed the automatic irrigation system is line insulators. Innov. Elec. Comm. Engg Proc. Fifth
ICIECE 2016, 311–317.
the subject of a review study.
Li, Kunlong. (2020). WITHDRAWN: Gymnastics training
This device, which is based on the RF module, is
action recognition based on machine learning and
used to send or receive radio signals between two wireless sensors. Microprocessors and Microsystems
devices. It has a complicated design due to the sensi- Available online 24 November 2020, 103522. doi:
tivity of radio circuits and the precision of the com- [Link] With-
ponents. A proposal was made by Ganai et al. (2022). drawn Article
A rain gun pipe with one end connected to the water Singh, Saravjeet, Jaiteg Singh, and Sukhjit Singh Sehra.
pump and the other to the plant’s root is used in this (2020). Genetic-inspired map matching algorithm for
sensor-based autonomous irrigation system with IoT. real-time GPS trajectories. Arabian Journal for Science
It does not use a sprinkler to deliver water and instead and Engineering. 45(4): 2587–2603.
relies on a soil moisture sensor. Anand et al. (2015) Mallikarjuna, B., Gulshan, S., and Meenakshi, S. (2022).
Blockchain technology: A DNN token-based ap-
demonstrated an Arduino-based IoT-enabled smart
proach in healthcare and COVID-19 to generate ex-
irrigation system. The researcher used an Arduino
tracted data. Exp. Sys. 39(3), e12778.
controller instead of a Raspberry Pi and did not use Mallikarjuna, B. (2020). Feedback-based fuzzy resource
soil moisture sensors. management in IoT-based-cloud. Int. J. Fog Comput.
(IJFC), 3(1), 1–21.
References Mallikarjuna, B. (2022). Feedback-based resource utili-
zation for smart home automation in fog assistance
Holliday, V. T. (1990). Methods of soil analysis, part 1, phys- IoT-based cloud. Res. Anthol. Cross-Dis. Des. Appl.
ical and mineralogical methods. Agron. Monograp., Automat., 803–824.
9(1), 87–89. Mallikarjuna, B., Viswanathan, R., and Bharat, B. N.
Knight, J. H. (1992). Sensitivity of time domain reflectom- (2019). Feedback-based gait identification using deep
etry measurements to lateral variations in soil water neural network classification. J. Crit. Rev. 7(4), 2020.
content. Water Res. Res., 28(9), 2345–2352. Mallikarjuna, B. (2022). The effective tasks management of
Marthaler, H. P., Vogelsanger, W., Richard, F., and Wieren- workflows inspired by NIM-game strategy in smart
ga, P. J. (2023). A pressure transducer for field tensi- grid environment. Int. J. Power Ener. Conver., 13(1),
ometers. Soil Sci. Soc. Am. J., 47(4), 624–627. 24–47.
Attema, E., Pierre, B., Peter, E., Guido, L., Svein, L., Ludwig, Mallikarjuna, B. (2022). An effective management of sched-
M., Betlem, R.-T., et al. (2007). Sentinel-1-the radar uling-tasks by using MPP and MAP in smart grid. Int.
mission for GMES operational land and sea services. J. Power Ener. Conver., 13(1), 99–116.
ESA Bul., 131, 10–17. Srinivasan, R., Mallikarjuna, B., Kavitha, M., Kavitha,
Gutiérrez, J., Juan, F. V.-M., Alejandra, N.-G., and Miguel, R., and Baharat, B. N. (2022). A comparative study:
Á. P.-G. (2013). Automated irrigation system using Wireless technologies in internet of things. 2022 2nd
a wireless sensor network and GPRS module. IEEE Int. Conf. Adv. Comput. Innov. Technol. Engg. (ICA-
Trans. Instrumen. Meas., 63(1), 166–176. CITE), 675–679.
Bhatia, S., Zainul, A. J., Shabana, M., and Neha, G. (2023). Mallikarjuna, B., Supriya, A., and Anusha, D. J. (2022). An
Integration of WSN and IoT: Wireless networks ar- improved deep learning algorithm for diabetes predic-
chitecture and protocols–A way to smart agriculture. tion. Handbook Res. Adv. Data Anal. Comp. Comm.
Handbook Res. Mac. Learn-Enabled IoT Smart Appl. Netw., 103–119.
Across Indust., 435–455. Brar, Preetinder Singh, Babar Shah, Jaiteg Singh, Farman
Sande, S. M. and Sharad, D. P. (2021). Controlling the Ali, and Daehan Kwak. (2022). Using modified tech-
growth of sugarcane plant in the nursery during germi- nology acceptance model to evaluate the adoption of
nation process by detecting and changing temperature a proposed IoT-based indoor disaster management
and humidity through IoT - A review. International software tool by rescue workers. Sensors. 22(5): 1866,
Journal of Research in Engineering and Technology, pp. 1–15.
8, 2395–0056. Mallikarjuna, B., Sathish, K., Gitanjali, J., and Venkata
Khattar, Nagma, Jaiteg Singh, and Jagpreet Sidhu. (2020). Krishna, P. (2021). An efficient vote casting system
An energy efficient and adaptive threshold VM con- with Aadhar verification through blockchain. Int. J.
solidation framework for cloud environment. Wireless Sys. Sys. Engg., 11(3–4), 237–256.
Personal Communications. 113, 349–367. Singh, A., Mallikarjuna, B., Mohammad, M., and Vaibhav,
Anand, K., Jayakumar, C., Mohana, M., and Sridhar, A. T. (2021). Design and implementation of superstick
(2015). Automatic drip irrigation system using fuzzy for blind people using internet of things. 2021 3rd Int.
logic and mobile technology. 2015 IEEE Technol. In- Conf. Adv. Comput. Comm. Con. Netw. (ICAC3N),
nov. ICT Agricul. Rural Dev. (TIAR), 54–58. 691–695.
Applied Data Science and Smart Systems 201
Bhatia, S., Mallikarjuna, B., Devraj, G., Urvashi, G., Suren- Bhatia, S., Zainul, A. J., and Shabana, M. (2023). A compar-
der, K., and Soniya, V. (2023). The future IoT: The cur- ative study of wireless communication protocols for
rent generation 5G and next generation 6G and 7G use in smart farming framework development. 2023
technologies. 2023 Int. Conf. Device Intel. Comput. 3rd Int. Conf. Intel. Comm. Computat. Tech. (ICCT),
Comm. Technol. (DICCT), 212–217. 1–7.
Ganai, P. T., Akash, B., Anita, S., Khairul, H. A., Sandeep, Bhatia, S., Zainul, A. J., and Shabana, M. (2023). Develop-
B., and Bhasker, P. (2022). A detailed investigation of ment and analysis of IoT based smart agriculture sys-
implementation of internet of things (IOT) in cyber tem for heterogenous nodes. 2023 Int. Conf. Recent
security in healthcare sector. 2022 2nd Int. Conf. Adv. Adv. Elec. Elect. Dig. Healthcare Technol. (REED-
Comput. Innov. Technol. Engg. (ICACITE), 1571– CON), 62–67.
1575.
28 Drowsiness detection in drivers: A machine learning
approach using hough circle classification algorithm for
eye retina images
J. Viji Gripsy1,a, N. A. Sheela Selvakumari2, S. Sahul Hameed3 and
M. Jamila Begam4
1
Department of Computer Science, PSGR Krishnammal College for Women, Coimbatore, Tamil nadu, India
2
Department of Computer Science, Sri Krishna Arts and Science College, Coimbatore, Tamil nadu, India
3
Department of Information Technology, Syed Hameedha Arts and Science College, Kilakarai, Tamil nadu, India
4
Department of Computer Science, Syed Hameedha Arts and Science College, Kilakarai, Tamil nadu, India
Abstract
Driving has become one of the most important routine works in our everyday life. For many people it is difficult to imagine
a life without driving. Accidents are a persistent and inevitable part of driving. Hence automatic drowsiness detection has
become a major challenge in research perspective. In this research work, drowsiness detection technique has been imple-
mented using machine learning (ML) techniques. In this methodology, a preprocessing, segmentation, feature extraction and
classification steps to perform. This work proposed hough circle (HC) classification algorithm for detecting drowsiness of the
eye retina images. The primary objective of this study is to evaluate the performance of the suggested hierarchical clustering
method through the utilization of diverse metrics. According to the results of the performance evaluation, the suggested HC
algorithm demonstrated a 90.8% accuracy rate, along with a minimal execution time and a lower error rate compared to
existing algorithms.
Keywords: Drowsiness detection, machine learning (ML), segmentation, feature extraction, classification, NB, SVM, k-NN
vijigripsy@[Link]
a
Applied Data Science and Smart Systems 203
Drowsiness detection
Driving is the riskiest job because while driving the
• Image indexing: image indexing method is used physical and psychological position must be paying
to expand the spatial position data. Spatial infor- attention. When there is required for attentiveness,
mation is stored in the rational databases. awareness, and non-sleepiness will not cause danger
• Database integration: The profitable dependency to the driver life’s (Sri Mounika et al., 2022). Although
normally generates 2 to 3 images of a specified of this sleepiness there are numerous reasons for
position every day and the expanded positions of motor vehicle crashes such as climate situation, path
each picture. situation, motor vehicle state, or less of driving ability.
• Spatial clustering: The cluster make of each spot There are three kinds of the category are been declare
is obtained after apply the clustering method. drowsiness that is joint restiveness, non-quick eye
• Semantic cluster concept generation: Semantic progress, and quick eye progress. Sleepiness acts as
cluster concepts such as middle cluster, left clus- the middle awakens is led to serious accidents. The
ter, intense cluster, sparse cluster, big cluster and mistake occurs when the person is drowsy or ignorant
small cluster. of the environment and not capable to make a perfect
• Trends and patterns mining: These trend and pat- decision on the path before colliding damage of the
tern are helpful for enhanced understanding of motor vehicle (Hardeep Singh, 2011).
the performance of the patterns mining (Hamzah The main idea behind this concept is to prevent
Al Najada, 2016). the drowsiness of the driver. Sleepiness is caused by
extended time driving which directs to highway area
Applications of image mining accidents (Gomathy et al., 2022). Up to 78% of car
Image mining is an upcoming development area, accidents are caused by sleepiness. The main cause for
while it is a latest research area, its expansion shown this is the lack of drowsiness chaos and being physi-
huge process. Image mining is using different area cally tired 20% is appropriate to drink and drive and
like medical, biometric, object or image detection 2% is of speed driving (Kundinger, Sofra, and Riener,
(Hamzah Al Najada, 2016). 2020). They are two kinds of identification approved
that is driver’s vehicle finding and driver’s facial recog-
Image mining in medical: Medical information is nition. To evade sleepiness the driver’s sequent check
using in our everyday life like C-T image sets, cardio- of the eye is essential; since in the extended journey,
gram image sets, MRI image sets, mammogram image the driver needs to position them self on the seat, if
sets, X-ray image sets and ultrasound image sets for not driver is patient through the driving and also with
finding problems. the authority and control the sleepiness (Moujahid
Image mining in biometric: Biometric informa- et al., 2021). Figure 28.2 illustrates the external eye
tion is using in our everyday life like school image structure.
sets, military image sets, hospital image sets, and There are complexities in the range of sleepi-
Image database evaluation mining feature extraction ness-connected accidents in the present day is no
204 Drowsiness detection in drivers
easy-to-predict, dependable approach for an exami- been observed but still, there is a problem with the
nation to decide whether sleepiness is an issue in the large vehicle (Rasna and Smithamol 2021). Reducing
accidents’ and, point of sleepiness the drivers’ physi- sleeping through crashes is a serious term of damage
cal and mentally painful. Most sleepy driving is a “not and loss of lives. There is an enlarged in the growth of
completely attentive” or “drowsy driving while he/she the detection system using the automotive application
is exhausted”. Numerically, the driver’s sleepiness is to make your fear of this difficulty.
Related works
Segmentation
This work used E_GRUNS a predictable algorithm
for segmenting the eye retina picture. Originally, the
picture center region is separated into three blocks:
left, center, and right. A comparatively huge center
region is measured in organize to take into explana-
tion cases where the retina is a little shift up or down
while passable field definition is unmoving maintain.
The E_GRUNS algorithms analysis the picture 0
inspects simply these three-block base on the follow-
Figure 28.3 System architecture ing two explanations:
Figure 28.3 illustrates system architecture • The retina has a huge figure of over control eye
beginning it. Thus, the block’s clear visible area
Pre-processing purpose is predictable to have the maximum in-
The pre-processing algorithm is used only for attri- formation of the three measured block drowsy
bute division because of eye images. This work used level stage.
Applied Data Science and Smart Systems 205
• The retina picture is noticeable by its high con- in the classification process. Furthermore, the per-
centration to generate segmentation. Hence, the formance of the eye drowsy decision is compared to
block containing the curve area is predictable conclusion.
to have the stage in order to find better drowsy This work proposed hough circle (HC) algorithm
detection three block drowsy levels (Gera et al., for classification. The proposed HC algorithm evalu-
2021; Gomathy et al., 2022). ates retina picture clearness and satisfies superiority
issues like clear visible area in the eye. The HC algo-
This work used the gray scale image as changes into rithm classification is full of eye retina images exposed
the optimized technique segmentation. Eye image is and using supervised classification. This research ret-
easy to find the drowsiness line curve and is also help- ina drowsy method is specially fit for retina pictures
ful for the next stage to find and concluded drowsi- to find sleepiness. The planned drowsy HC algorithm
ness using the black and white region of the eye. is calculated as the most of the retina drowsy value in
Segmentation was implemented to guarantee the ret- white and black region space color. The conclusion
ina picture clear visible area information during the of the HC algorithm creates an improved suitable
different procedures. Segmentation checks and inner method (Kundinger, Sofra, and Riener, 2020). Based
quality ratio (IQR) were used. The image connected on the HC algorithm, to find sleepiness is calculated
to the overall and standard point of every picture as the retina region for the classification to give better
boundary, correspondingly. results.
To exact classification the retina picture estimate
Feature extraction decision as it reveals the information content inside
Feature extraction reduces the amount of informa- the picture information. The picture of every stage is
tion that must be processed, while still precisely and calculated as the standard of the left, right, up, and
telling the unique data set. Feature extraction natural down as given by the following equations.
measures like sleepy eye are extremely entity and dif-
fer from one person/subject to the next person/sub-
ject uniqueness for further analysis. The procedures
used in this research are obtained from the drowsi-
where the white region images have larger in the non-
ness detection dataset. This dataset includes non-
drowsy. The planned retina picture estimate decision
physiologic information, acquire in an investigational
as it reveals the information contained inside the pic-
field study in a real lab. This work used the FBSG
ture information. The picture of every stage is calcu-
algorithm to extract each driver’s sleepiness stage is
lated as the standard of the left, right, up, and down
also accessible in the dataset (Dua et al., 2021). The
as given by the following equations.
original set of unrefined data is reduced to a further
convenient group for the process. A feature of this
outsized dataset is a huge number of variables that
need a lot of computing resources for the procedure.
Feature extraction is the forename for a method that
selects and/or combines variables into feature extrac- Algorithm for HC
tion, efficiently the total of records that should be
processed, to improve precision and entire unique Multi-stage range of R-curve
datasets. Step 1. Circle Selection: Fit a circle to the locate of
the person bend bi and bright pixels in The
dark state
Classification
Classification of sleepy recognition issues in eye pic- Step 2. Select the bend lie in the environs of this
circle for the point
tures was the major cause of identifying the problem.
Details and non-finding problems inside the retina Step 3. Superior Selection: the bend into two
categories: a) black, b) white
picture result due to means will not proceed to the
next step so it needs an achievable algorithm method Step 4. Calculate the close relative vessel-segment
direction bθi For each bend bi
to find the drowsiness in the classification. Albadawi,
Takruri, and Awad (2022) observed process recogni- Step 5. Examine each picture in steps of 90°
tion method and proper classification retina picture Step 6. if circle/curve (s) exist then
process are helpful to lead the anther stage to find the Step 7. if bi
drowsiness to guide the classification result (Hardeep Step 8. is right then
Singh, 2011). The proposed algorithm is introduc- Step 9. if numerous then
ing that method based on the classification process.
Step 10. choose r- bend with the least bend angle
Numerous studies and comparisons were performed
206 Drowsiness detection in drivers
Algorithm for HC
Step 11. else
Step 12. choose r- bend
Step 13. end if
Step 14. end if
Step 15. else
Step 16. carry on
Step 17. end
Dataset
The approach is illustrated through drowsiness detec-
tion, utilizing a carefully selected dataset derived from
the MRL eye dataset. This dataset serves as a special-
ized subset designed for efficient categorization tasks.
It encompasses a diverse range of infrared images of
human eyes, encompassing both low and high-quality
photographs obtained under varying lighting condi-
tions and using different imaging devices. This dataset
is ideally suited for evaluating a multitude of features
or trainable classification models. To facilitate algo-
rithmic comparisons, the images are categorized into
multiple classes, making them highly suitable for both
training and testing classification algorithms. In total,
the dataset comprises 216 photographs, with 194
depicting non-sleepy subjects and 22 depicting indi-
viduals experiencing drowsiness.
Performance measures
Various performance metrics are employed to thor-
oughly assess the effectiveness of both the proposed
and existing algorithms. The evaluation of the sug-
gested algorithm’s performance encompasses an Figure 28.5 Pre-processing result using noise reduc-
array of comprehensive measures, including PSNR tion, sharpness, contrast, brightness
(Peak Signal-to-Noise Ratio), recall, mean absolute
error (MAE), precision, execution time, accuracy
and F-score. These metrics provide a holistic and in- the brightness provides better pre-processing results
depth examination of the algorithm’s performance in among the other approaches in this work.
diverse aspects, ensuring a robust assessment of its
capabilities. Segmentation
Figure 28.6 depicts segmentation comparisons for
Pre-processing right eye image thresholds 1.5, 1.6, 1.8, and image
Figure 28.5 represents the PSNR result graph using threshold 1.7 using the E_GRUNS algorithm.
noise reduction, sharpness, contrast, and brightness
of the right eye image for noise reduction, sharpness, Performance of the classifier
contrast, and brightness. Figure 28.4 shows the pre- Based on the HC outcome, the attribute-derived con-
processing comparison between proposed pre-pro- clusion is to fill with better performance. In the seg-
cessing images. From this analysis, it is observed that ment, every different image attribute set is estimated
Applied Data Science and Smart Systems 207
Department of Computer Science, PSG College of Arts and Science, Coimbatore, Tamil nadu, India
2
Department of Computer Science (PG), PSGR Krishnammal College for Women Coimbatore, Tamil nadu, India
3
Abstract
Wireless sensor networks (WSNs) offered promising opportunities for the development of ubiquitous and pervasive comput-
ing. However, the implementation of WSNs encountered many barriers and problems. These included the dynamic nature of
network topology and the occurrence of congestion, both of which had a detrimental impact on network capacity utilization
and overall performance. In WSN systems, the transmission of packets occurs from nodes with low congestion levels to nodes
with high congestion levels, resulting in a decrease in energy levels for nodes located in close proximity to the sink nodes. The
effective rate control with data aggregation (ERCDA) strategy employs an efficient data aggregation technique to enhance
the equitable utilization of battery power across all nodes involved. The proposed methodology is executed on the NS2.35
platform and evaluated in terms of throughput, packet loss, end-to-end delay, and source data transmission rate adjustment.
According to simulations, it has been shown that the use of ERCDA exhibits a higher degree of efficacy in comparison to
conventional congestion-handling approaches.
Keywords: Wireless sensor networks, congestion control, data aggregation, rate control, ERCD
a
[Link]@[Link]
210 Optimizing congestion collision using effective rate control
the originals. The total bytes of all packets to transmit to decrease latency. Artificial neural network
are packet size. (ANN) training and testing for high-bandwidth
traffic of variable burstiness. Time(p,q) is packet
(1) flow time, k(p,q) is packet transmission, n(p,q) is
packet count, BWreq(p,q) is desired bandwidth,
(2) and DR is DataRate.
Existing node nz € Backward (a, df2) while n2 € techniques. This phenomenon occurs as a result of the
Ns(m1) L m1) For(e, fi) or nz € For(e, fi). allocation of priority levels to traffic classes at each
} virtual queue and the equitable distribution of band-
Step 12: Eliminate Coding collision width across all nodes in the network.
// Training using ANN
Step 13: xp,g = {k(p.q). n(p.a). a(p.q). BWreq(p.q). Packet loss
duration(p.q)} It is the amount of data dropped or missed during
Step 14: duration(p+1.q) transfer.
Step 15: y=ANN (xp,q)
(7)
Simulation results
In this section, the ERCDA technique is executed Figure 29.2 illustrates the percentage of packet
in network simulator version 2.35 (NS2.35) and loss for the HOCA, DRCDC, CCR, and ERCDA
1000×1000m2 simulation area with 50 nodes, 5GHz approaches over different simulation durations,
operating frequency, and 120 s simulation time. Its measured in seconds. The findings suggest that the
effectiveness is analyzed compared to the Health- ERCDA approach has a lower incidence of packet
care-aware Optimized Congestion Avoidance and loss in comparison to other strategies. If the simula-
Control protocol (HOCA), Differentiated Rate tion duration is 120 s, the ERCDA algorithm has a
Control Data Collection (DRCDC), and Congestion- packet loss rate of 20%, which is the lowest among
aware Clustering and routing (CCR) techniques. The the other methods. Therefore, the ERCDA exhibits
analysis is conducted based on throughput, packet little packet loss as a result of its implementation of
loss, end-to-end (E2E) delay, and source data transfer virtual queues and equitable allocation of bandwidth
rate adjustment. among nodes to effectively manage congestion inside
the WSN.
Throughput
It is the amount of data accepted by the target within End-to-end delay
a time.
It is the time taken for a data to be broadcasted from
an origin to the sink.
(6)
Conclusion
This research introduces the ERCDA approach, which
takes into account factors such as energy usage, bat-
tery power, and power management. Network cod-
ing is used in situations when the data rate exceeds
a predetermined threshold value, taking into account
Figure 29.3 E2E delay the stated threshold value for the data rate. To opti-
mize the equitable utilization of battery power, a
proficient approach is implemented, including the
aggregation of data, coding conditions, and coding
collision mechanisms. In conclusion, the simulation
results demonstrate that the efficacy of the ERCDA
approach is superior to that of traditional congestion
management strategies.
References
Ahmad Jan, M., Roohullah Jan, S., Usman, M., and Alam,
M. (2018). State-of-the-art congestion control pro-
tocols in WSN: A survey. EAI Endor. Trans. Internet
Things, 3(11),154379. [Link]
3-2018.154379.
Figure 29.4 Data transfer rate Batra, U. Annual IEEE Computer Conference, IEEE Interna-
tional Advance Computing Conference 4 2014.02.21-
22 Gurgaon, International Advanced Computing
It is observed that the ERCDA approach has lower Conference 4 2014.02.21-22 Gurgaon, and IACC 4
end-to-end latency in comparison to other strategies. 2014.02.21-22 Gurgaon. n.d. IEEE Int. [Link].
If the duration of the simulation is set to 120 s, the Conf.(IACC), 2014 21-22 Feb. 2014, Gurgaon, India.
end-to-end delay of the ERCDA is measured to be Cheng, J., Ye, Q., Jiang, H., Wang, D., and Wang, C. (2013).
STCDG: An efficient data gathering algorithm based
161 ms, which is comparatively lower than the delays
on matrix completion for wireless sensor networks.
seen in other techniques. Hence, the least E2E is cor-
IEEE [Link].,12(2), 850–861. https://
related with the greatest throughput and reduced [Link]/10.1109/TWC.2012.121412.120148.
packet loss. Farsi, Mohammed, Mahmoud Badawy, Mona Moustafa,
Hesham Arafat Ali, and Yousry Abdulazeem. (2019).
Data transfer rate adjustment A congestion-aware clustering and routing (CCR)
It is the data transfer rate of origin, which handles the protocol for mitigating congestion in WSN. IEEE Ac-
congestion and buffer overflow in WSN. cess. 7: 105402–105419. pp. 1–18.
The data transmission rate (measured in packets per Ghaffari, Ali. (2015). Congestion control mechanisms in
second) for the HOCA, DRCDC, CCR, and ERCDA wireless sensor networks: A survey. Journal of net-
approaches is shown in Figure 29.4. The simulation work and computer applications. 52, 101–115.
Kafi, M. A., Ben-Othman, J., Ouadjaout, A., Bagaa, M.,
period (measured in seconds) is varied to observe the
and Badache, N. (2017). REFIACC: Reliable, efficient,
performance of these techniques. The findings of this
fair and interference-Aware congestion control proto-
investigation suggest that the ERCDA approach dem- col for wireless sensor networks. Comp. Comm.,101,
onstrates superior data transfer rates as a result of its 1–11. [Link]
efficient rate adjustment and effective allocation of Kafi, M. A., Djenouri, D., Ben-Othman, J., and Badache,
bandwidth. If the duration of the simulation is 120 s, N. (2014). Congestion control protocols in wire-
it can be seen that the data rate of ERCDA is 57 pack- less sensor networks: A survey. IEEE [Link].
ets per second, which surpasses the data rates of other Tutor.,16(3),1369–1390. [Link]
methods. The ERCDA has the capability to progres- SURV.2014.021714.00123.
sively decrease the data transmission rate in relation
214 Optimizing congestion collision using effective rate control
Mazunga, Felix, and Action Nechibvute. (2021). Ultra-low work for cloud [Link]. Per. Comm.,113,
power techniques in energy harvesting wireless sen- 349–367.
sor networks: Recent advances and issues. Scientific Tshiningayamwe, L., Lusilao-Zodi, G. A., and Dlodlo, M.
African. 11: e00720. doi: [Link] E. (2016). A priority rate-based routing protocol for
sciaf.2021.e00720 wireless multimedia sensor networks. [Link].
Monowar, M. and Bajaber, F. (2017). Towards differenti- Comput., 419, 347–358. [Link]
ated rate control for congestion and hotspot avoid- 3-319-27400-3_31.
ance in implantable wireless body area networks. UC Santa Cruz. UC Santa Cruz Electronic Theses and Dis-
IEEE Acc.,5,10209–10221. [Link] sertations. Adaptive Network Coding In MANET.
ACCESS.2017.2708760. (n.d.). [Link]
Rezaee, A. A., Hossein Yaghmaee, M., and Masoud Rah- Wang, Chonggang, Kazem Sohraby, Bo Li, Mahmoud
mani, A.(2014). Optimized congestion management Daneshmand, and Yueming Hu. (2006). A survey
protocol for healthcare wireless sensor networks. Wire. of transport protocols for wireless sensor networks.
Per. Comm.,75(1),11–34. [Link] IEEE network. 20(3): 34–40.
s11277-013-1337-z. Wang, F., Liu, W., Wang, T., Zhao, M., Xie, M., Song, H., Li,
Singh, J., Goyal, G., andGill, R. (2020). Use of neuromet- X., and Liu, A. (2019). To reduce delay, energy con-
rics to choose optimal advertisement method for sumption and collision through optimization duty-
omnichannel [Link]. Inform. Sys.,14(2), cycle and size of forwarding node set in WSNs. IEEE
243–265, [Link] Acc.,7,55983–55915. [Link]
40392. CESS.2019.2913885.
Sarode, Sambhaji S., and Jagdish W. Bakal. (2018). A data Xie, K., Wang, L., Wang, X., Xie, G., and Wen, J.(2018).
transmission protocol for wireless sensor networks: Low cost and high accuracy data gathering in WSNs
A priority approach. Journal of Telecommunication, with matrix completion. IEEE [Link]-
Electronic and Computer Engineering (JTEC). 10(3): put.,17(7),1595–1608. [Link]
65–73. TMC.2017.2775230.
Sumathi, R. and Srinivasan, R. (2012). QoS aware routing Yaakob, N. and Khalil, I. (2016). A novel congestion
protocol to improve reliability for prioritisedheteroge- avoidance technique for simultaneous real-time
neous traffic in wireless sensor network. Int.J. Paral. medical data transmission. IEEE J. Biomed. Health
[Link]. Sys.,27(2),143–168. [Link] Informat.,20(2),669–681. [Link]
0.1080/17445760.2011.608356. JBHI.2015.2406884.
Swain, S. K. and Pradipta Kumar, N. (2019). Priority Yin, Long, Jinsong Gui, and Zhiwen Zeng. (2019). Im-
based adaptive rate control in wireless sensor net- proving energy efficiency of multimedia content
works: A difference of differential approach. IEEE dissemination by adaptive clustering and D2D mul-
Acc.,7,112435–11247. [Link] ticast. Mobile Information Systems. [Link]
CESS.2019.2935025. org/10.1155/2019/5298508.
Tan, J., Liu, W., Wang, T., Zhang, S., Liu, A., Xie, M., Ma, M., Zhuang, Y., Yu, L., Shen, H., Kolodzey, W., Iri, N., Caulfield,
and Zhao, M. (2019). An efficient information maxi- G., and He, S. (2019). Data collection with accuracy-
mization based adaptive congestion control scheme in aware congestion control in sensor networks. IEEE
wireless sensor [Link] Acc.,7,64878–64896. [Link].,18(5), 1068–1082. [Link]
[Link] org/10.1109/TMC.2018.2853159.
Khattar, N., Singh, J., andSidhu, J. (2020). An energy effi-
cient and adaptive threshold VM consolidation frame-
30 DDoS attack detection methods, challenges and
opportunities: A survey
Jaspreet Kaura and Gurjit Singh Bhathal
Department of Computer Science and Engineering, Punjabi University, Patiala, India
Abstract
A concept labeled“cloud computing”enables online, on-demand access to a shared pool of computing resources. It offers
scalability, flexibility, and cost-efficiency to organizations, allowing them to focus on their core business functions. However,
the popularity and widespread adoption of cloud computing also make it an attractive target for cyber-attacks. Distributed
denial of service (DDoS) is one such threat, where a network or service is overwhelmed with an excessive amount of mali-
cious traffic, rendering it inaccessible to legitimate users. DDoS attacks may affect cloud-based services’performance and
availability, resulting in severe loss of revenue and damaging their reputation. To combat these attacks, various classification
methods have been developed to categorize DDoS attacks based on their characteristics and attack vectors. These classi-
fications aid in understanding attack patterns, developing effective defense mechanisms, and enhancing incident response
strategies. Furthermore, attacks using DDoS are often planned to utilizebotnets, which are networks of compromised com-
puters under the direction of one administrator. Botnets amplify the impact of DDoS attacks by harnessing the combined
resources of multiple compromised devices. Understanding the operation and behavior of botnets is crucial for mitigating
DDoS attacks effectively. This paper provides a description of cloud computing, a look at attacks via DDoS, a method to
categorize them, and the role of botnets in carrying out these attacks focusing on the significance of having strong security
mechanisms set up in cloud systems in order to protect against these types of risks.
a
jasspreet87270@[Link]
216 DDoS attack detection methods, challenges and opportunities
accessible. It’s distributed nature has made it very easy temperature of healthcare workers as and when
to attack (Alanazi etal.,2019). Here are a few cloud required that too any hassle of removing hand gloves
computing applications or services that can help you or PPE kits. Another objective is to make this face
fulfill your corporate, academic, and general business shields reusable (Figure 30.1).
goals (Kati etal., 2020).
DDoSattacks and their classifications
The architecture of cloud computing
Several types of security attacks can threaten the secu-
The structure and features that make up a cloud com- rity of cloud computing environments. Attacks such
puting system are known as the cloud computing as DDoS create a serious security risk for cloud com-
architecture. It typically involves multiple layers of puting infrastructures. A DDoS attack involves the use
abstraction and various hardware and software com- of multiple computers or other devices to bombard
ponents, which are written below. a targeted system with requests or traffic, overload-
Physical infrastructure layer– This layer consists of ing it and keeping it unreachable to authorized users.
physical servers, storage devices, networking equip- Cloud-based services are especially vulnerable to
ment, anddata centers that provide the foundation for DDoS attacks because they rely on the internet to con-
cloud computing. nect with users and other services, which makes them
Virtualization layer– This layer allows the ability to more susceptible to traffic floods and network conges-
generate and keep track of virtual machines (VMs), tion. Strong security measures like firewalls, intrusion
and virtual resources, such as virtual CPUs, memory, detection and prevention systems, and other tools
and storage. It allows multiple VMs to run on maxi- should be considered by organizations, and content
mizing the use of hardware resources. delivery networks (CDNs) to protect against attacks
Platform layer–Developers may design and deploy involving DDoS in the cloud. They should also regu-
their software on this layer without having to be con- larly monitor their network traffic and be prepared
cerned about managing the underlying infrastructure. to respond quickly to any suspended or confirmed
Application development, testing, deployment, and DDoS attacks. Host may include virtual machines,
scalability tools are included. computers, laptops, or zombies in this attack. They
Application layer–Applications and services that come with a remote. Multiple computers are used in a
are provided to end users via the internet are included threatening attack, or DDoS (Abusaimeh etal., 2020)
in this layer. It includes cloud computing services. (Figure 30.2).
Management layer–This layer provides tools and
services for monitoring, managing, and securing the Classification of DDoSattack
cloud computing environment. It includes tools for
managing VMs, storage, networking, security, and DDoS attacks are very difficult to stop or identify since
compliance services. they spread easily over the internet (Kesavamoorthy
This paper is aimed to design a technologically etal., 2020). There are some different classifica-
advanced 3D face shield capable of monitoring body tions based on threats such as attacks on lower rate
requests the infected machine sends fewer requests at ICMP floods –Internet control message protocol
a higher session rate than normal users, and the ses- (ICMP) floods send a number of packets to the tar-
sion rate can fluctuate at random. In lower session geted network or server by taking advantage of vul-
attacks per second, a breached computer utilizes the nerabilities in the ICMP protocol.
requests at higher rates but at a lower session rate UDP floods –UDP floods are similar to volumetric
than usual, and the session rate can change at random attacks, but they target the UDP protocol, which is
(Kumar etal., 2016). The two categories of DDoS are, used for real-time applications such as video stream-
firstly, a fraudster uses an attack known as DDoS that ing and online gaming.
consumes a lot of bandwidth to flood network equip-
ment, bandwidth allocations, and communication Based on application level flooding attacks
channels with a huge number of packets. Second,in
application-layer DDoS attacks an attacker hijacks HTTP floods –A lot of HTTP requests are overload-
the computing resources of the computer web host- ing the web server. The volumetric attack is unique to
ing the victim of the attack and prevents it from pro- spoofing techniques (Dong etal., 2019).
cessing genuine transactions and requests by taking SIP flood attack –When used for communication,
advantage of the behavior of services and applications voice over IP (VoIP) uses SIP for call signaling. SIP
as well as that of computer communication protocols telephones can effectively be overloaded with mes-
(TCP, HTTP, etc.). sages, making it impossible for them to deal with
valid requests (Dong etal., 2019).
Reflective/amplified attacks –Reflected DDoS or
Based on the attack traffic characteristics DRDoS attacks operate at the application level using
Volumetric attacks – The most common DDoS attack TCP, UDP, or a combination of the two (Kshirsagar
is a volumetric attack. These attacks aim to take over etal., 2021). In these attacks, the attacker requests
a network or server with a flood of traffic, usually servers vulnerable to overflow of amplified traffic on
using botnets (networks of compromised devices) to the target system, causing the target system to crash.
generate a high volume of traffic. TCP-based resembled DDoS attacks include those
that use Microsoft SQL server (MSSQL) and the
Protocol attacks simple service discovery protocol (SSDP). Character
TCP SYN floods –TCP SYN floods exploit the TCP generator protocol (CharGen) attacks and simple
protocol’s three-way handshake process to consume file transfer protocol (TFTP) attacks are examples of
server resources and prevent legitimate traffic from UDP-based directed DDoS (Kshirsagar etal., 2021).
reaching the server. CGI request attack – The victim’s computer stops
DNS –Domain names must be turned into IP responding to requests as a result of the attacker’s
addresses in cloud-based settings with the domain large number of CGI requests utilizing the victim’s
name system (DNS). Attacks affect the DNS infra- computer’s CPU resources (Dong etal., 2019).
structure that is DDoS-related and has an effect on Slowloris attacks–Slowloris attacks are a type
the accessibility of cloud services. Examples include of DDoS attack that exploits vulnerabilities in web
DNS query floods or DNS amplification attacks that server software by sending a high number of slow
overwhelm DNS servers or deplete their resources. HTTP requests to consume server resources.
218 DDoS attack detection methods, challenges and opportunities
Fraggle attack –The Fraggle attack is a particular that can then be used to launch DDoS attacks on a
kind of DDoS attack that sends a lot of UDP traffic to server (Sanjeetha etal.,2021; Brar et al., 2022).This
the switch’s transmission organization. This is similar section provides an overview of the botnet’s archi-
to a Smurf attack that uses UDP rather than ICMP tecture and the tools used to perform DDoS flooding
(Mohmand etal., 2022). attacks. If an attacker uses botnets or zombies, devel-
oping an effective and efficient defense mechanism
LDDoS (low-rate) and HDDoS (high-rate) becomes more difficult. This is primarily due to two
causes. First, many zombies would be participating in
High-rate DDoS (HDDoS) and low-rate DDoS the attacks to increase their size and disruptiveness.
(LDDoS) (Daffu etal., 2017), which are often referred Second, it is clear that zombies operating under the
to as brute force and semantic attacks (Liu etal., 2019), attacker’s command use faked IP addresses, making
respectively, are two categories into which denial of it very difficult to track them back (Kesavamoorthy
service (DDoS) attacks typically fall. The high-rate etal., 2020) (Figure 30.3).
attack is capable of a volume of more than 500 Gbps IRC Botnet –Most bots are useful and safe and essen-
and attempts to either prevent actual users from con- tial to the performance of the internet. The earliest
necting, which is referred to as a recourse depletion internet bots enabled classroom activity using the inter-
attack is a type of bandwidth depletion attack that net relay chat (IRC) protocol (Wainwright etal., 2019;
make cloud services unavailable (Radain etal., 2021). Singh et al., 2020). Some operates of the IRC botnet.
Attackers use a brute-force attack, also known as a Compromising devices –Are often operated via an
flooding attack or a high-rate DDoS attack, to send IRC botnet. The attacker uses malware or vulnerabili-
enormous malicious requests that completely absorb ties to infect a significant amount of devices, including
the network capacity of the targeted cloud server computers, servers, and IoT devices.
(Alanazi etal., 2019).On the other hand, vulner- IRC communication –On the infected devices, the
abilities in protocols can be accessed using semantic bot virus creates a connection to an IRC server.
attacks, also known as vulnerability attacks, rather Command and management –Using an IRC server,
than by consuming any available network or cloud the botnet operator gives orders and directives to the
computing resources. To specifically target a certain bots.
protocol or application, the attacker creates a small Activity on a botnet –The bots execute the given
amount of malicious traffic. These types of attacks actions after receiving instructions from the IRC
are also referred to as low-rate attacks on distributed server.
denial of service. Low-rate attack traffic approaches Update andmaintaining –The IRC botnet can com-
legal traffic in appearance. As a result, low-rate DDoS municate with the botnet operator via the IRC server
attacks consequently are harder to identify than high- to get updates or new instructions.
rate DDoS attacks(Alanazi etal., 2019). Web/HTTP botnet –A drive-by download, spam,
and other similar methods are typically used to down-
Based on botnet-attack load an HTTP-based bot at first. Consequently, a par-
ticular constant-size transmission between an infected
Botnets can be used as part of a DDoS attack. host and a new target is [Link] botnets only
Malware can be installed on servers to build botnets communicate with their C&C servers sometimes to
obtain requests, rather than maintain a continuous Mohmand, M.I., Hameed, H., Ali Khan, A., Ullah, U., Za-
connection (Letteri etal., 2020).Web-based bots can karya, M., Ahmed, A., Raza, M., Izaz Ur, R., and Hal-
be set up and run and managed via complicated PHP eem, M. (2022). A machine learning-based classifica-
scripts, which also implement encryption over the tion and prediction technique for DDoS attacks. IEEE
Acc.,10, 21443–21454.
HTTP (port 80) or HTTPS (port 443) protocols for
Wainwright, Polly, and Houssain Kettani. (2019). An analysis
communication (Kesavamoorthy etal., 2020).
of botnet models. In Proceedings of the 2019 3rd Inter-
national Conference on Compute and Data Analysis.
Conclusion 116–121. [Link]
Daffu, P. and Kaur, A. (2016). Mitigation of DDoS attacks
The basic information about cloud computing and in cloud computing. 2016 5th Int. Conf. Wire. Netw.
its layer-based design, including the risks of DDoS Embed. Sys. (WECON), 1–5.
attacks, was covered in this paper. DDoS attacks, Liu, X., Jiadong, R., He, H., Wang, Q., and Song, C. (2021).
which try to overload a target system by flooding it Low-rate DDoS attacks detection method using data
with overwhelming traffic or resource requests, are compression and behavior divergence measurement.
well-known as an extreme risk to online services and Comp. Sec., 100, 102107.
networks. DDoS attacks can be categorized according Dong, Shi, Khushnood Abbas, and Raj Jain. (2019). A sur-
to the characteristics of the attack flow, application- vey on distributed denial of service (DDoS) attacks in
level flooding attacks, and Botnet-based DDoS attacks SDN and cloud computing environments. IEEE Ac-
cess. 7: 80813–80828.
are one particular kind of DDoS attack. In these
Radain, D., Almalki, S., Alsaadi, H., and Salama, S. (2021).
attacks, an attacker is in control of a botnet, which is a
A review on defense mechanisms against distributed
network of compromised devices. By commanding the denial of service (ddos) attacks on cloud comput-
bots to flood the target system with traffic, the attacker ing. 2021 Int. Conf. Women Data Sci. Taif University
may instruct the botnet to perform DDoS attacks. (WiDSTaif), 1–6.
Organizations need to have strong security measures Kshirsagar, Deepak, and Sandeep Kumar. (2021). A feature
set up if they want to reduce the risks imposed by the reduction based reflected and exploited DDoS attacks
context of cloud computing, and attacks using DDoS. detection system. Journal of Ambient Intelligence and
Humanized Computing. 13, 393–405.
Alanazi, S.T., Anbar, M., Karuppayah, S., Al-Ani, A. K.,
References and Sanjalawe, Y. K.(2019). Detection techniques for
Abusaimeh, Hesham. (2020). Distributed denial of service DDoS attacks in cloud environment. Intel. Interact.
attacks in cloud computing. International Journal of Comput. Proc. IIC 2018, 337–354.
Advanced Computer Science and Applications, 11(6). Agrawal, Neha, and Shashikala Tapaswi. (2019). Defense
1–6. mechanisms against DDoS attacks in a cloud comput-
Kati, Sherwin, Abhishek Ove, Bhavana Gotipamul, Mayur ing environment: State-of-the-art and research chal-
Kodche, and Swati Jaiswal. (2020). Comprehensive lenges. IEEE Communications Surveys & Tutorials.
Overview of DDOS Attack in Cloud Computing En- 21(4), 3769–3795.
vironment using different Machine Learning Tech- Singh, S., Singh, J., and Sehra, S. S. (2020). Genetic-inspired
niques. Available at SSRN 4096388. Proceedings of map matching algorithm for real-time GPS trajecto-
the International Conference on Innovative Comput- ries. Arab. J. Sci. Engg., 45(4), 2587–2603.
ing & Communication (ICICC) 2022, 1–14. Sanjeetha, R., Raj, A., Saivenu, K., Ahmed, M. I., Sathvik,
Kesavamoorthy, R., Alaguvathana, P., Suganya, R., and B., and Kanavalli, A. (2021). Detection and mitigation
Vigneshwaran, P. (2020). Classification of DDoS at- of botnet based DDoS attacks using catboost machine
tacks–A survey. Test Engg. Manag., 83, 12926–12932. learning algorithm in SDN environment. Int. J. Adv.
Kumar, V. and Kumar, K. (2016). Classification of DDoS Technol. Engg. Explor. 8(76), 445.
attack tools and its handling techniques and strategy Wainwright, P. and Kettani, H. (2019). An analysis of bot-
at application layer. 2016 2nd Int. Conf. Adv. Comput. net models. Proc. 2019 3rd Int. Conf. Comput. Data
Comm. Automat. (ICACCA)(Fall), 1–6. Anal., 116–121.
Dong, S., Khushnood,A., and Raj, J. (2019). A survey on Letteri, I., Giuseppe, D.P., and Pasquale, C. (2019). Fea-
distributed denial of service (DDoS) attacks in SDN ture selection strategies for HTTP botnet traffic de-
and cloud computing environments. IEEE Acc. 7, tection. 2019 IEEE Eur. Symp. Sec. Priv. Workshops
80813–80828. (EuroS&PW), 202–210.
Brar, P. S., Shah, B., Singh, J., Ali, F., and Kwak, D. (2022). Abusaimeh, Hesham. (2020). Distributed denial of service
Using modified technology acceptance model to evalu- attacks in cloud computing. International Journal of
ate the adoption of a proposed IoT-based indoor di- Advanced Computer Science and Applications, 11,
saster management software tool by rescue workers. 1–6. DOI:10.14569/ijacsa.2020.0110621
Sensors, 22(5), 1866. Agrawal, N. and Tapaswi, S. (2019). Defense mechanisms
Kshirsagar, D. and Kumar, S. (2022). A feature reduction against DDoS attacks in a cloud computing environ-
based reflected and exploited DDoS attacks detection ment: State-of-the-art and research challenges. IEEE
system. J. Amb. Intel. Human. Comput., 1–13. Comm. Sur. Tutor., 21(4), 3769–3795.
31 A review of privacy-preserving machine learning
algorithms and systems
Utsav Mehtaa, Jay Vekariya, Meet Mehta, Hargeet Kaur and Yogesh Kumar
Department of CSE, School of Technology, Pandit Deendayal Energy University, Gandhinagar, Gujarat, India
Abstract
There is an enormous amount of data being gathered, and great scope of information-based learning is possible through
this data, termed machine learning (ML). Along with this, there is a larger concern for the privacy of the users whose data is
being gathered and used for model training purposes. This recent scenario has set the base for advancement in the research
field of privacy preserving in machine learning (PPML). This technique focuses on designing systems and algorithms that
can perform information-based learning, keeping the data of the user protected. This paper reviews and provides concise
information on the several algorithms and systems proposed to achieve the goal of PPML. A background of works pertain-
ing to learning on both the type of outsourced as well as distributed data systems are covered in this paper followed by a
description of the proposed algorithms that aim to preserve privacy and are modifications to existing algorithms like SVM,
kNN and neural networks are presented in this paper.
utsavmmehta17@[Link]
a
Applied Data Science and Smart Systems 221
each model has not memorized the data but rather end server that performs computing on the obtained
has learned from the general pattern. data. This work proposes the usage of a homomorphic
In order to make the vote of each learner private a encryption technique for sharing the data securely
Laplace noise is added to the votes, the Laplace noise between the users and performing computations
is given as Equation 2. using encrypted data. A general representation of the
encrypted form of the data is shown in Equation 6.
This paper uses fully homomorphic encryption (FHE)
for the encryption similar to what is mentioned in
(Zvika, 2014) and (Dan, 2005).
Papernot et al. (2018) have improved the PATE in
the order it scale capable over large size of data. This
work modifies the system for the original PATE archi-
tecture by changing the noisy max computation. where,
The classical Laplace noise max is given as shown D* = User data attributes
in Equation 3 pk* = Public key
Following the pre-processing, this work imple-
mented three protocols under the POS system,
. kernel matrix protocol, SVM model setup, SVM
classification.
The Gaussian noise max is given as shown in the Samanthula et al. (Bharath et al., 2014) have
Equation 4 proposed a kNN method for semantically secure
encrypted outsourced data, termed PPkNN. This
. focuses on the security related to the usage of user A’s
data for identifying user B’s output class using kNN.
The above-mentioned works majorly focuses on This work proposes the architecture of the kNN algo-
the approach that is general for the privacy-preserv- rithm in the outsourced data using two cloud serv-
ing. However, the following discussed works are spe- ers C1 and C2 which are semi-honest, in such a way
cific in nature for distributed and outsourced data that neither the data of user A nor the query and
(Cynthia, 2008). class labels of user B are revealed to each other. This
Hwanjo et al. (2006) developed an SVM that paper proposes two schemes SSkNN and SCMCk for
focuses on privacy preserving in the distributed data achieving the goal of neighbor identification and class
model. Conventionally in the linear kernel to draw prediction in a secure manner.
separating boundaries between the data points. The Ping et al. (2017) have proposed a PPDL model.
individual data points are used for the derivation of This work again focuses on a PPML through an
the kernel using the kernel matrix. This work pro- outsourced data system. This work proposed an
poses the development of SVM with non-linear ker- advanced schema for cloud computations using neu-
nels using the gram matrix given in Equation 5 ral networks. The pre-processing using double encryp-
tion techniques using BCP (Emmanuel, 2003) and
MK-FHE. Adriana et al. (2012) set up on two plat-
. forms, one is the cloud C and the other is an autho-
rized center AU. The users are needed to upload the
Secure set intersection cardinality was used for data, by encrypting using the BCP scheme, followed
gram matrix computation as suggested in Vaidya et by blinding of the same through C, the same blinded
al. (Jaydeep et al., 2005). The equation proposed in cipher-text is shared to AU which contains the master
this work is focused on horizontally separated data. key to decrypt the cipher-text and then re-encrypt it
Following the work for non-linear kernels for hori- using MK-FHE before sending it back to C. The cloud
zontally partitioned data, Yunhong et al. (2009) pro- plays the role to compute the model using a neural
posed modifications in the gram matrix in order to network on the re-encrypted data.
develop both linear and non-linear kernels for ver-
tically separated data in multiparty or distributed
systems. Analysis and results
Fang et al. (2015) have proposed the SVM with Abadi et al. (2016) have proposed a DP-SGD presents
PPML in outsourced data systems. The term used a range of notable advantages, including pioneering
for this system in this work is POS (protocol for out- algorithmic approaches for training deep neural net-
sourced SVM). This work proposed a mechanism for works under privacy constraints, which stands as a sig-
the chain of supply of data between the users and the nificant contribution to the field of privacy-preserving
Applied Data Science and Smart Systems 223
machine learning. Furthermore, the paper introduces measures like encryption and decryption, which may
a refined framework for assessing privacy costs using increase processing time and resource requirements.
differential privacy, enhancing our ability to ensure In terms of applications, this technique holds prom-
data privacy in model training. These innovations ise in various domains that demand secure data han-
have versatile applications, particularly in image dling, including medical diagnosis, financial analysis,
classification and language representation, offering and fraud detection. It enables a wide range of pri-
enhanced privacy without sacrificing model utility. vacy-conscious data mining and ML tasks, ensuring
Beyond these domains, the techniques introduced in that sensitive data remains protected while valuable
the paper hold promise for safeguarding sensitive insights are extracted.
information in various sectors like healthcare, finance, The technique discussed in the Vaidya et al. (2005)
and government, where large datasets are prevalent. offers several significant advantages. Firstly, it excels
Moreover, they open doors to privacy-preserving ML in preserving data source privacy, enabling data min-
solutions in critical areas such as recommendation ing while rigorously protecting sensitive information.
systems, fraud detection, and autonomous vehicles, Additionally, it boasts proven security properties,
where data privacy remains paramount. However, ensuring the confidentiality of data during the min-
it’s important to acknowledge the inherent trade-off ing process. The paper introduces efficient protocols
between privacy protection and model quality, and for generating association rules from disparate parties
the need for rigorous evaluation and risk assessment holding private information about the same individu-
in sensitive applications due to the complexity of als, facilitating collaborative data mining while pre-
interpreting deep neural network representations and serving privacy. Furthermore, it presents a vision of
potential fine-grained data encoding. a versatile toolkit for privacy-preserving data mining
Papernot et al. (2018) proposed PATE approach approaches, potentially enhancing the adaptability of
offers significant advantages in the realm of pri- privacy-conscious data mining techniques. However,
vacy-preserving ML. One of its notable strengths there are notable limitations to consider. The tech-
is its ability to enhance privacy protection through nique may be susceptible to collusion among parties,
the introduction of selective and less noisy aggrega- which could compromise data privacy protections. It
tion mechanisms for teachers’ answers. This robust may not be applicable in fully malicious settings where
approach provides strong differential-privacy guar- some parties actively attempt to undermine privacy
antees, effectively safeguarding individual privacy in safeguards. Maintaining security properties when
scenarios involving sensitive data. However, there are combining different secure computations, especially
potential drawbacks to consider, such as the possibil- in iterative data mining scenarios, can be challenging.
ity of utility trade-offs. Like many privacy-preserving The paper also highlights an open issue related to re-
techniques, PATE may lead to a reduction in the accu- running algorithms after minor data changes, which
racy of the student model, potentially necessitating could impact its practicality. In terms of applications,
increased computational resources and longer train- the technique demonstrates its practicality in privacy-
ing times. This limitation could be a concern in appli- preserving association rule mining by securely com-
cations where model accuracy is critical. Nonetheless, puting the intersection cardinality of distributed sets.
the broad applicability of the PATE approach in tasks It also hints at potential uses in constructing decision
involving sensitive data, such as personal messages or trees, EM clustering, and managing association rules
medical records, is a significant advantage. It enables in vertically partitioned data between two parties,
the extraction of valuable insights while preserving highlighting its potential as a foundational compo-
individual data privacy and improving model accu- nent for creating a comprehensive toolkit supporting
racy through innovative aggregation mechanisms. various privacy-conscious data mining techniques.
The technique discussed by Hwanjo et al. (2006) The privacy-preserving outsourced SVM (POS)
offers several notable advantages. Firstly, it excels technique (Fang et al. 2015) offers a robust solution
in the preservation of data privacy while still deliv- for safeguarding individual privacy while enabling
ering accurate results, making it invaluable in situ- collaborative operations on encrypted and outsourced
ations where privacy and security concerns are data. It’s versatility shines through in its capability to
paramount. Secondly, it is particularly effective for maintain data privacy across various data partition-
scenarios involving horizontally partitioned data, ing schemes, encompassing horizontal, vertical, and
facilitating distributed knowledge discovery while arbitrary divisions, making it adaptable to a wide
upholding robust privacy protection. However, there array of scenarios. However, it’s worth noting that the
are some potential drawbacks to consider. One sig- paper does not explicitly address all potential privacy
nificant limitation is its relative computational inef- concerns, particularly those related to securely pro-
ficiency when compared to traditional SVM methods. cessing and storing encrypted and outsourced data
This inefficiency arises from the additional privacy in cloud environments, leaving room for additional
224 A review of privacy-preserving machine learning algorithms and systems
privacy considerations beyond its scope. In terms scenarios involving multiple data owners with dif-
of applications, POS proves particularly valuable in ferent datasets. One specific application highlighted
the realm of secure support vector machine (SVM) in the paper is privacy-preserving face recognition, a
classification on outsourced and encrypted data. Its prominent biometric authentication technique used in
versatility extends its utility across diverse domains, real-life scenarios. Furthermore, the generic nature of
including healthcare, finance, and social networks, the proposed solutions allows for their application in
where preserving data privacy is of paramount impor- various other ML tasks that share the same privacy-
tance. Furthermore, it complements other privacy- preserving setting and requirements.
preserving SVM methods, especially in the context
of distributed models. This emphasis on maintaining
Conclusion and future scope
user data confidentiality, integrity, and auditability
within cloud environments enhances privacy protec- The paper provides a concise review of several works
tion in collaborative data analysis and classification and research done in the field of PPML to the readers.
tasks, offering a promising approach to secure and Several works focusing on privacy concerns in differ-
privacy-conscious ML. ent environments of data flow are presented in their
Samanthula et al. (2014) introduces a k-NN pro- paper. Modifications into several existing algorithms
tocol with several notable advent ages. Firstly, it like SVM, kNN, and neural networks are also high-
provides robust data confidentiality and privacy pro- lighted. A few techniques based on DP were discussed
tection, ensuring the security of user input queries. that focused on the privacy of data while the training
Additionally, the protocol incurs negligible compu- of the model, followed works that focused on the pri-
tation costs on the end-user, enhancing its efficiency vacy of distributed and outsourced data.
and user-friendliness. However, there are several chal- These models however developed and proposed,
lenges and concerns to consider. The protocol differs there is still a scope for them to make easy to use for
from existing privacy-preserving classification tech- several consultancy firms as well as several general
niques by hosting encrypted data on the cloud, poten- freelancing users in order to achieve the goal of PPML.
tially introducing unique operational challenges. A part of it can be achieved by making libraries in
While it aims to address accuracy issues that can arise commonly used programming languages like python
in existing methods due to the introduction of statis- in popular ML libraries like sckit-learn. The exist-
tical noise, it may also face challenges in mitigating ing works focus on modifying existing algorithms to
data access pattern leakage. In terms of applications, achieve privacy, however, a lot of room exists for the
the proposed k-NN protocol holds relevance in data development of specific models having the capability
mining contexts, with specific use cases mentioned in of achieving the goal of privacy.
fraud detection in the financial sector and tumor cell
level prediction in healthcare. This suggests practical
References
applications in domains where data privacy and clas-
sification accuracy are paramount considerations. Abadi, M., Andy, C., Ian, G., Brendan McMahan, H., Ilya,
Ping et al. (2017) introduces multi-key privacy- M., Kunal, T., and Li, Z. (2016). Deep learning with
preserving deep learning schemes with several nota- differential privacy. Proc. 2016 ACM SIGSAC Conf.
Comp. Comm. Sec., 308–318.
ble advantages. These schemes effectively safeguard
Dan, B., Eu-Jin, G., and Kobbi, N. (2005). Evaluating
sensitive data, intermediate results, and the training 2-DNF formulas on ciphertexts. Theory Cryptograp.
model, ensuring robust privacy measures. A security Second Theory Cryptograp. Conf., TCC 2005, Cam-
analysis included in the paper further validates the bridge, MA, USA, February 10-12, 2005. Proceedings
effectiveness of these techniques, assuring users of 2, 325–341.
their data’s confidentiality. Moreover, their versatil- Zvika, B., Gentry, C., and Vaikuntanathan, V. (2014). Leveled
ity and adaptability make them suitable for a wide fully homomorphic encryption without bootstrapping.
range of ML tasks within the same privacy-preserv- ACM Trans. Comput. Theory (TOCT), 6(3), 1–36.
ing setting. However, there are certain limitations Emmanuel, B., Catalano, D., and Pointcheval, D. (2003). A
to consider. Implementing these schemes may incur simple public-key cryptosystem with a double trap-
additional computational and communication costs door decryption mechanism and its applications. Int.
Conf. Theory Appl. Cryptol. Inform. Sec., 37–54.
compared to traditional deep learning methods due
Bing, C. and Titterington, D. M. (1994). Neural networks: A
to encryption and decryption processes, potentially
review from a statistical perspective. Statist. Sci. 2–30.
impacting efficiency. Additionally, the dependency Tom, D. (1995). Overfitting and under computing in ma-
on trusted third parties for secure key management chine learning. ACM Comput. Sur. (CSUR), 27(3),
could pose practical challenges in some scenarios. In 326–327.
terms of applications, the schemes find valuable use in Cynthia, D. (2008). Differential privacy: A survey of results.
cloud computing, especially in collaborative learning Int. Conf. Theory appl. Models Comput., 1–19.
Applied Data Science and Smart Systems 225
Ehsan, H., Hassan, T., Mehdi, G., and Wright, R. N. (2018). Nicolas, P., McDaniel, P., Sinha, A., and Wellman, M. P.
Privacy-preserving machine learning as a service. Proc. (2018). Sok: Security and privacy in machine learn-
Priv. Enhanc. Technol., 2018(3), 123–142. ing. 2018 IEEE Eur. Symp. Sec. Priv. (EuroS&P),
Jordan, M. I. and Mitchell, T. M. (2015). Machine learn- 399–414.
ing: Trends, perspectives, and prospects. Science, Papernot, Nicolas, Shuang Song, Ilya Mironov, Ananth Rag-
349(6245), 255–260. hunathan, Kunal Talwar, and Úlfar Erlingsson. (2018).
Arya, R., Singh, J., and Kumar, A. (2021). A survey of mul- Scalable private learning with pate. arXiv preprint
tidisciplinary domains contributing to affective com- arXiv:1802.08908, 1–34. [Link]
puting. Comp. Sci. Rev., 40, 100399. arXiv.1802.08908
Ping, L., Li, J., Huang, Z., Li, T., Chong-Zhi, G., Siu-Ming, Samanthula, B. K., Elmehdwi, Y., and Jiang, W. (2014). K-
Y., and Kai, C. (2017). Multi-key privacy-preserving nearest neighbor classification over semantically se-
deep learning in cloud computing. Fut. Gen. Comp. cure encrypted relational data. IEEE Trans. Knowl.
Sys., 74, 76–85. Data Engg., 27(5), 1261–1273.
Fang, L., Ng, W. K., and Zhang, W. (2015). Encrypted SVM Vaidya, J. and Chris, C. (2005). Secure set intersection car-
for outsourced data mining. 2015 IEEE 8th Int. Conf. dinality with application to association rule mining. J.
Cloud Comput., 1085–1092. Comp. Sec., 13(4), 593–622.
Adriana, L.-A., Tromer, E., and Vaikuntanathan, V. Vargas, Rocio, Amir Mosavi, and Ramon Ruiz. (2017).
(2012). On-the-fly multiparty computation on the Deep learning: a review. Queensland University of
cloud via multikey fully homomorphic encryption. Technology, Creative Commons Attribution 4.0, 1–11.
Proc. 44th Annual ACM Symp. Theory Comput., Jingting, X., Xu, C., and Bai, L. (2019). DStore: A distrib-
1219–1234. uted system for outsourced data storage and retrieval.
Arvind, N. and Vitaly, S. (2008). Robust de-anonymization Fut. Gen. Comp. Sys., 99, 106–114.
of large sparse datasets. 2008 IEEE Symp. Sec. Priv. Hwanjo, Y., Jiang, X., and Vaidya, J. (2006). Privacy-pre-
(sp 2008), 111–125. serving SVM using nonlinear kernels on horizontally
Papernot, Nicolas, Martín Abadi, Ulfar Erlingsson, partitioned data. Proc 2006 ACM Symp. Appl. Com-
Ian Goodfellow, and Kunal Talwar. (2016). Semi- put., 603–610.
supervised knowledge transfer for deep learning Hu, Y., Liang, F., and Guoping, H. (2009). Privacy-preserv-
from private training data. arXiv preprint arX- ing SVM classification on vertically partitioned data
iv:1610.05755, 1–16. [Link] without secure multi-party computation. 2009 5th
arXiv.1610.05755 Int. Conf. Nat. Comput., 1, 543–546.
32 Optimization techniques for wireless body area network
routing protocols: Analysis and comparison
Swati Goel, Kalpna Guleriaa and Surya Narayan Panda
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
Wireless body area networks (WBANs) are a type of wireless network used to monitor and collect data from various sensors
attached to the human body. WBAN has enormous applications like assisted living, healthcare, sports, defense, entertain-
ment, military, and many others. The successful deployment and operation of WBANs come with several challenges, includ-
ing energy efficiency, on-time data transmission, reliability, scalability, security, and network lifetime. Routing is a critical
and important aspect of WBAN communication, and several optimization techniques have been proposed to improve the
routing performance in WBANs. Optimization techniques plays a crucial role in optimizing and improving the performance
of WBANs routing protocol. These techniques aim to optimize various parameters such as power consumption, data rate,
energy efficiency, throughput, network bandwidth, delay, network lifetime etc. This paper provides an overview of the op-
timization techniques used for routing in WBANs in recent years (2014–2023). These optimization techniques can signifi-
cantly improve the performance of WBAN routing protocols and enable them to be used for various applications related to
medical and non-medical domains. The paper discusses the various optimization techniques proposed in the literature, their
strengths, weaknesses and simulators used to evaluate WBAN routing performance. By analyzing and summarizing existing
research, this paper aims to provide valuable insights into the current state of optimization techniques in WBAN routing and
identify potential research gaps for future exploration.
Keywords: Wireless body area networks, technology, optimization techniques, sensors, WBAN, routing, energy
[Link]@[Link]
a
Applied Data Science and Smart Systems 227
(Guleria, Kumar, and Verma, 2019; Bhola et al., 2022). Article organization
As technology continues to advance, ongoing research The paper’s organization is shown with the help of the
and development in optimization will further improve road map provided below as shown in Figure 32.1.
the performance and capabilities of WBANs, ultimately Initially, a brief introduction about WBAN, WBAN
benefiting both healthcare providers and patients. routing, and the need for optimization followed by
research questions are defined which is named as an
Objective of the paper introduction followed by the methodology adopted
To be used as a foundation for future study and the for article selection. Following this the background
creation of more sophisticated WBAN routing proto- is discussed and provides answers to our first two
cols, the goal of this work is to provide insight into research questions RQ1 and RQ2 which corresponds
the optimization approaches currently employed in to our RQ3 and provides state-of-the-art work related
WBAN routing with the help of a systematic review to WBAN routing using optimization techniques. A
and to outline the various advantages and limitations comparative study of various WBAN routing proto-
of the various optimization techniques used to per- cols based on optimization techniques highlighting
form efficient WBAN in recent last decade i.e., from the main objective, advantages, disadvantages, and
2014 to 2023. simulation tool used. Finally the paper is concluded
with a focus on future work.
Contribution
In this article; a systematic study is conducted to pres- Methodology
ent the current state of art of the optimization tech-
niques used in WBAN routing. The importance of The study deals with the work done in the field of
optimization algorithms for WBAN routing protocols WBAN routing using optimization techniques for the
is highlighted. The recent published research papers years 2014–2023. The different research questions
between the years 2014–2023 for routing protocols are framed to provide a systematic review which is
using optimization techniques are presented in this termed as RQ1, RQ2, RQ3, and RQ4. The articles
review article. The systematic review methodology is were searched from various databases including
followed to select the high-quality research articles to Scopus, Google Scholar, Dimensions, Web of Science
justify the work. Four research questions (RQs) are (WoS), IEEE Xplore, and ScienceDirect. The aim of
framed and addressed to conduct the review in an effi- the survey is to help the researchers to develop new
cient way for better understanding and to organize routing algorithms using optimization techniques
the review. The research questions are defined below by working on the limitations of the existing rout-
named as RQ1 to RQ4. ing protocols in future. The review was done in vari-
ous stages which is illustrated in Figure 32.2. The
RQ1. What is the annual trend of growth in the articles were selected having effective information
WBAN domain? in the domain of optimization techniques in WBAN
RQ2. How is the routing associated and important routing. Various combinations used for searching
in WBAN? the articles from the databases were – “Wireless
Body Area Network or WBAN and routing proto-
RQ3. Which state-of-the-art optimization techniques cols”, “Optimization techniques in WBAN”, and
are used in the WBAN domain to perform efficient “Wireless Body Area Network or WBAN and opti-
routing? mization techniques and routing protocols”. Over
RQ4. What are the limitations and challenges involved 1065 articles were shown for the above-mentioned
in WBAN routing protocols? search strings. Thirty-three articles were removed due
Table 32.1 Inclusion and exclusion criteria. selected for further refinement. Again, the exclusion
of the articles was done based on the abstract, con-
Criteria Description
clusion, and methodology resulting in 31 articles.
Inclusion Papers from recent years (2014–2023) Finally, 18 articles that concentrates on high quality
research work in the field of WBAN routing using
Papers that focused on recent research
trend, opportunities and challenges related optimization techniques were shortlisted that corre-
to WBAN routing using optimization sponds to our review criteria. Figure 32.2 shows the
techniques systematic structure followed to refine the articles:
Papers written in English language only Table 32.1 provides an insight on the inclusion and
Papers that mention the clear methodology exclusion criteria on the basis of which the articles
were selected and eliminated to be included in the
Papers from reputed journals and
conferences study:
Exclusion Duplicate papers/identical titles
Papers in which methodology is not
Background
present or unclear WBANs contains sensors attached to the human body
Papers that are not open-access to monitor various human body vital parameters like
Papers that do not entail WBAN heart rate, sugar-level, pH level, BP, temperature, etc.
optimization techniques as primary study The sensors communicate with a central node called
as sink or sink node, which aggregates the data and
sends it further to base station. Medical WBANs pro-
to duplications resulting in 1032 articles for the next vide an evolutionary shift from illness to wellness,
step. After applying the exclusion criteria based on with a focus on early identification and detection of
the year (2014–2023), language (English only), and disease which can reduce global healthcare expendi-
title of research 748 papers were excluded. Then ture by over $400 trillion annually. The global BAN
the articles from the reputed journals and confer- market growth rate is predicted to grow at a CAGR of
ences having state-of-art techniques were analyzed 22.3% for the year 2022–2032. It is estimated to be
and segregated from the remaining articles. On this valued at about US$229.8 Bn by 2032, going up from
basis; 107 articles were found to be relevant and was US$30.8 Bn in 2022.
Applied Data Science and Smart Systems 229
The data packets in WBANs are small in size and and GA together for choosing the optimal routing.
have a limited transmission range. Therefore, the rout- GA is used to generate the initial pheromone distribu-
ing of these data packets in WBANs is a challenging tion and ACA is used to convert pheromones distribu-
task due to their limited transmission range and small tion into pheromones; positive feedback from ACA is
packet size. Energy-efficient routing, temperature- used for finding the optimal solution. It has reduced
based routing, cluster routing, QoS-aware routing the energy consumption according to the simulation
and cross-layer optimization are some of the routing results for every sensor node. The main limitation is
categories that are widely used in WBANs (Kour and that it has not considered node degree and multi-path
Kang, 2019) (Bhatia, Panda, and Nagpal, 2020). The routing which can be part of future work to enhance
aim of the routing protocols in WBAN is to increase the overall performance.
efficiency, improve reliability, enhance security, mini- A cluster-based energy-efficient optimization algo-
mize delay, increase throughput, enhance stability, rithm called modified ANT colony was proposed by
and network lifetime making them suitable for vari- the authors Rakhee and Srinivas (2016) to find the
ous applications, particularly healthcare. Routing next hop in an optimal way using probabilistic func-
protocols in WBANs facilitate the intelligent and tion based on residual energy and pheromones in each
optimal routing of data packets among sensors and node. OMNet++ simulator is used by the authors for
sink nodes. Various routing protocols were designed the implementation and the result shows that the
in the past based on traditional routing approaches proposed system performed better when compared
to enhance the overall network performance (Abidi, in terms of latency, energy, jitter, and throughput. It
Jilbab, and Mohamed, 2020; Goyal et al., 2023). But helps to choose the optimal path for data delivery in
nowadays in the past few years; optimization tech- indoor environments of the hospitals for continuous
niques come into existence that help to optimize the monitoring of patients in BAN. The proposed algo-
routing algorithms to provide a better and enhanced rithm ensures better network connectivity using a
performance as compared to traditional routing breadth-first search algorithm as it uses a modified
approaches. Numerous studies have been published CH rotation process.
in the past that use optimization to increase the effec- Kaur and Singh (2017), in their work introduced
tiveness of energy use, power consumption, reliabil- a multi-objective cost function for selecting the for-
ity, congestion and QoS requirements for a WBAN. warder node which is optimized using a GA. A reli-
This study includes a review paper for optimization able and energy-efficient routing have been proposed
techniques used in WBAN to perform efficient rout- to process the important data based on optimal cri-
ing such as genetic algorithms (GA), fuzzy-logic, bio- teria. The forwarder node is chosen based on the
inspired techniques, etc. that enhances lifetime of cost function’s minimum value. Except energy con-
networks in terms of duty cycle and QoS criteria. The sumption, reliability model, and path-loss model,
need for optimization techniques to enhance routing the proposed work offers GA-based optimization to
efficiency WBANs is paramount due to several key perform efficient routing. The proposed model per-
reasons that contribute to enhanced network perfor- formed better and consumed less energy due to the
mance as these optimization techniques support vary- use of multi-hop communication. The network model
ing signal strengths; provide higher energy efficiency, can be expanded in future work to take into account
support mobility, and control temperature rises by more complex network circumstances, including the
enabling seamless communication and optimized various network topologies accountability and cross-
routing (Seth, Panda, and Guleria, 2021). layer interactions.
To obtain less energy usage and shorten transmis-
sion times, the suggested framework optimizes the
Review on optimization techniques used in
shortest path at various phases of data collection. In a
WBAN routing
study by Ali and Al Masud (2018), the bees algorithm
In the related work, the various papers that uses opti- is employed as an optimization technique to help
mization techniques such as cuckoo search optimiza- with WBAN deployment and improve WBAN trans-
tion, spider monkey optimization, ant lion algorithm, mission efficiency during Hajj. To overcome these
dragonfly optimization, lion optimization, grey wolf difficulties and identify the shortest path for data in
optimization, whale optimization algorithm, etc., the least amount of time during the congested Hajj
which are used in the literature by the researchers to environment, the bees algorithm is used. The bees
perform the efficient routing have been presented. The algorithm exhibits good performance in MATLAB
review consists of the latest studies of the last 10 years simulations when it comes to lowering transmission
from 2014 to 2023. time, energy consumption, latency, and throughput.
The authors, Xu and Wang (2014) have used the Additionally, bees algorithm, employed as an opti-
hybrid approach by using ant colony algorithm (ACA) mization tool to choose the shortest path in several
230 Optimization techniques for wireless body area network routing protocols: Analysis and comparison
phases, ensures that data reaches its destination in making appropriate CH selections. One of the key
the shortest amount of time with the least amount of elements that contribute to a longer network lifespan
energy and delay. is the head node selection process, which takes into
The author’s main goal (Bilandi, Verma, and Dhir, account the energy that is still available. Additionally,
2019) is to implement a routing mechanism that it lowers the no. of duplicate packets sent and received,
uses the PSO optimization technique in conjunction conserving the entire network’s energy. Future energy
with relay node selection based on distances and optimization and node balancing techniques could
residual energy. According to findings, the suggested make use of the cuckoo search optimization (CSO)
protocol perfectly balances the need for fewer relay and grey wolf (GW) algorithms.
nodes with the energy-saving WBAN. The fundamen- WBAN’s constant monitoring and data transmis-
tal drawback of the study is that nodes are static in sion system secures the patient’s life. The most used
this and moreover the link is bi-directional. The sug- method in WBAN for load balancing is clustering that
gested approach increases the network’s lifetime and offers an effective approach for the energy optimiza-
improves the WBAN’s reliability. tion of sensor nodes. In the paper by Mehmood and
The authors, Panhwar et al. (2020) used the GA for Aadil (2021), the authors proposed an optimization
the selection of the best routing path. Unlike previ- technique for cluster formation using evolutionary
ously available direct techniques; this approach cal- algorithms. The authors recommend the dragonfly
culates the distance between the nodes under multiple method (DA) as the most effective method since it
scenarios while considering factors like the energy creates the fewest optimized, and long-lasting clus-
used by sensor nodes, the number and position of ters, hence extending the network lifetime where CHs
sensor nodes, the distance between the deployed sen- are chosen based on fitness value. The primary draw-
sors, and the number of rounds. The use of GA drops back of the study is that the temperature of the sensor
fewer packets as compared to a traditional approach. nodes in WBAN is not considered.
It also outperforms in terms of the dead nodes and Within a WBAN, a body node coordinator (BNC)
energy that increased the WBAN lifetime significantly. is in charge of organizing and receiving data transmis-
For future work the sensor nodes can be increased sions from bio-sensor nodes. Therefore, it is essential to
in number and the cloud can be used for storing the position BNC in WBAN in an ideal location to reduce
data. energy consumption of the network during data trans-
The author’s objective is to develop an energy opti- mission. This research (Choudhary, Nizamuddin, and
mization technique for inter-BAN communications Zadoo, 2022) presents a full data routing approach
based on evolutionary algorithm and cluster-based for WBAN that integrates a cluster routing protocol
routing. In work by Aadil et al. (2020); the authors with a multi-objective particle swarm optimization
proposed an optimization technique named GOA (MO-PSO) based BNC placement technique. The sug-
which is a metaheuristic approach to solve the opti- gested method makes use of a particle structure with
mization problem. The clustering process is optimized three variables, the first two of which give the BNC
in WBAN which helps to improve the overall network coordinates and the third of which specifies the CH
life as it consumes less energy. The shorter cluster for- node. The fitness of MO-PSO particles is estimated
mation means a high frequency of re-clustering those using a multi-objective fitness evaluation operator
results in high network computational cost and over- using the average bit error rate (BER) and network
head in communication. energy consumption. The model develops into an ideal
A metaheuristic method for choosing the best clus- BNC location that simultaneously reduces average
ters in WBANs was put out by the Saleem et al. (2021) BER and network energy consumption. Additionally,
to implement an energy-efficient routing protocol for a lower BER results in a significant boost to the net-
monitoring livestock behavior and health. The sug- work throughput rate. A network architecture that is
gested method uses ALO to choose the best clusters optimized by the proposed MO-PSO model establishes
for various pasturage sizes with various transmission the best position for BNCs. The reduced BER and net-
ranges while taking into account user preferences work energy consumption objectives are effectively
for cluster density. To assure the best CH selection met by the optimized network design. The network
in the livestock industry and to increase the lifetime throughput rate significantly increases as a result of
of WBANs, the proposed protocol provides a major the decreased BER those results.
contribution. The created network is said to have a An adaptive cuckoo search (ACS) algorithm is pre-
mesh topology; however, because the animals in this sented by Samal, Patra, and Kabat (2022) to minimize
network are dynamic, it is challenging to maintain network energy consumption and locate relay nodes.
this topology. In this, the number of optimal relay nodes placement
The whale optimization (WO) algorithm approach problem in the WBAN scenario is formulated. The
was suggested by Li and Jiang (2022) as a means of ACS is used for the selection of relay node that uses
Applied Data Science and Smart Systems 231
the fitness function considering energy consumption, and energy and the simulation results shows that the
coverage of sensors, distance, cost, etc. The proposed proposed system using ACO technique performs bet-
scheme suffers from relay locating problems, unreli- ter than conventional systems. The major limitation
able transmission, and direct transmission. of the work is that various important parameters
According to Arafat, Pan, and Bak (2023), the that can affect WBAN routing performance are not
authors presented a distributed routing protocol ignored like node temperature, reliability, delay, etc.
called DECR that is based on a two-hop method. It So, further enhancements can be made in the future
is based on a clustering process, where during the to overcome the problems occurring due to network
cluster formation phase, information about neighbor partitioning and topology change.
nodes is received within a two-hop range. For CH When compared to other routing protocols the sim-
selection and optimization, MGWO, a meta-heuristic ulation results shows that the use of GA optimization
optimization algorithm inspired by nature, is used. By makes the network more energy efficient. The packets
lowering the transmission distance to each CH, the dropped and dead nodes are parameters taken into
hierarchical grey-wolf optimization technique assists account while it ignored the cross-layer interactions
in data transfer based on residual energy and node and network topologies.
connection in each cluster. As it employs energy-effi- Ali and Al Masud (2018) considered various
cient clustering and an ideal CH selection process, the parameters such as transmission time, throughput,
network lifetime is increased. energy consumption, and delay based on the objec-
It can be analyzed from the literature that several tive to overcome many challenges faced by pilgrim
routing protocols are implemented using different during Hajj.
optimization techniques to optimize the various per- The authors Bilandi, Verma, and Dhir (2019) used
formance metrics to enhance overall WBAN’s per- PSO for selecting optimized relay node. The major
formance. The literature review helps to lay a strong parameters considered in the work are energy, dis-
foundation for the concepts, the optimization tech- tance of nodes, stability period and throughput.
nique, and the simulation tool used in existing work The use of GA by Panhwar et al. (2020) achieved
for optimizing the WBAN routing protocols. better PDR, lesser number of dead nodes, and reduced
energy consumption but in this path loss factor
Comparative analysis of routing protocols using parameter is ignored which can result in delayed data
optimization techniques in WBAN transmission.
Aadil et al. (2020) uses various parameters like
An analytical comparison of different optimization direction, density, speed, grid size for CH selection.
algorithms used in WBAN routing protocols was car- Saleem et al. (2021) considered energy and tempera-
ried out. A comparison table is made that highlights ture of the nodes for selecting the CH but they have
the objective of each routing protocol focusing on not considered the dynamic nature of WBAN’s.
the optimization technique used to achieve the speci- The proposed scheme (Ibrahim, 2021) used
fied objectives. The table also discusses about advan- Dragonfly optimization technique that enables forma-
tages, disadvantages and lists the simulation tool used tion of efficient clusters thus focusing on the network
to implement the routing protocol and the impact lifetime parameter.
of optimization techniques in WBAN routing. The Choudhary, Nizamuddin, and Zadoo (2022), the
authors provided the drawbacks of existing optimiza- parameters like energy efficiency, network throughput
tion algorithms so that it can help future researchers rate and bit error rate (BER) are considered to enhance
to design a more efficient routing protocol by elimi- the overall network performance. The ACS scheme is
nating the limitations of the existing routing protocols used to find the optimal set of relay nodes based on
by using efficient optimization techniques. The com- energy parameter. The proposed algorithm by Samal,
parison is carried out based on several QoS metrics Patra, and Kabat (2022) selects a set of relay nodes to
such as network stability, reliability, security, network make the optimal selection of cost, energy consump-
lifetime, throughput, packet delivery ratio, residual tion from the candidate sets considering the coverage
energy, no. of packets dropped/received, number of of sensors. Cost, energy consumption, coverage, and
dead nodes/alive nodes, end-to-end delay, etc. distance are the factors that are utilized to calculate
The different parameters considered by the authors the fitness function in the algorithm for selecting the
Xu and Wang (2014) to enhance the routing quality optimal number of relay nodes.
is by optimization of energy consumption focusing on The grey wolf optimization algorithm is used by
time and quality. But the parameters such as residual Arafat, Pan, and Bak (2023) to ensure energy-efficient
energy, multi-path routing, node degree are not consid- data packet delivery. The node connectivity and the
ered. The authors Rakhee and Srinivas (2016) consid- residual energy parameters are considered for select-
ered various parameters like jitter, latency, throughput ing the CH in each cluster.
Table 32.2 Comparative analysis of routing protocols using optimization techniques
232
S. Reference and Optimization Objective Simulation Advantages Limitations
Optimization techniques for wireless body area network routing protocols: Analysis and comparison
No. year technique used tool used
1 Xu and Wang, GACA (Genetic To enhance routing quality and - It combines the benefit of both GA Certain factors like residual
2014 Ant Colony prolong network lifetime by and AC thus improving the time energy, multi-path routing, and
Algorithm) optimizing energy consumption duration and quality of the network. node degree are not considered
It balances and reduces the energy
consumption to prolong the network
lifetime
2 Rakhee and ACO (Ant To choose an optimal path OMNeT++ The proposed system has better This system is designed and is
Srinivas, 2016 Colony for monitoring vital signs for performance in terms of latency, jitter, limited to indoor environments of
Optimization) continuous data delivery of energy, and throughput than the the hospital for data delivery. It
patients in indoor environments conventional systems can be further extended to work
in hospitals in outdoor environments
3 Kaur and Singh, GA (Genetic To perform energy-efficient MATLAB Selection of the best optimal routing Packets dropped and dead
2017 Algorithm) routing and selecting optimal path is done using GA and energy- nodes are only considered for
forwarder nodes using GA based efficient routing is provided in this performance evaluation. In
on a multi-objective cost function work future work, more complex
network scenarios like cross-layer
interactions and network topologies
can be taken into account
4 Ali and Al BA (Bees The objective is to optimize the MATLAB The use of bees algorithm to select In this work, only the energy
Masud, 2018 Algorithm) path by using bees optimization the shortest path in multiple phases consumption factor is taken
algorithm to achieve low energy helped to reduce the delay by enabling into account; other network-
consumption and reduced the data to reach the destination in the related QoS parameters are not
transmission time shortest possible time and reducing the considered
energy consumption significantly
5 Bilandi, Verma, PSO (Particle To develop an energy-efficient MATLAB The proposed protocol minimizes the Low throughput and poor
and Dhir, 2019 Swarm mechanism of routing that uses number of relay nodes for energy- network stability. Future research
Optimization) PSO with the relay node selection efficient WBAN. It performs better in can consider the analysis of
for heterogeneous WBAN terms of residual energy as compared multiple BANs and can focus on
to state-of-art protocols reliable and secure data delivery
6 Panhwar et al., GA (Genetic To select the best routing path MATLAB It saves the energy significantly of Pathloss is high. The number
2020 Algorithm) and to select the nearest node the WBAN to increase the network of sensors can be increased and
using GA optimization for lifetime. It also has a better PDR, a cloud data storage can be used to
calculating distances between lesser number of dead nodes, and enhance the performance in the
the nodes for energy-efficient reduced energy consumption which future
transmission significantly saves energy
7 Aadil et al., Goa Algorithm To develop an energy MATLAB Increase in network efficiency due The computational cost
2020 optimization technique for inter- to the use of intelligent and optimal is increased and also the
BAN communications based clustering techniques communication overhead due to
on evolutionary algorithms and shorter cluster lifetime resulting in
cluster-based routing high frequency of reclustering
S. Reference and Optimization Objective Simulation Advantages Limitations
No. year technique used tool used
8 Saleem et al., ALO (Ant Lion To design an energy-efficient MATLAB The temperature of the nodes remains The constructed network is
2021 Optimizer) routing protocol by selecting controlled due to the association of defined to be in mesh topology,
optimal clusters in WBANs for the limited number of nodes with less the animals in this network
livestock health and behavior energy for forming CH are dynamic and hence the
monitoring maintenance of this topology is
difficult
9 Li and Jiang, WO (Whale To intelligently select the CH to MATLAB The proposed scheme produces results The number of clusters is
2022 Optimization) increase network’s lifetime and with high accuracy. It consumes low predicted randomly which results
Algorithm reduce the duplicate packets thus energy and enhances network lifespan. in an unbalanced number of
saving the network’s energy It is also capable of finding the nodes in each cluster
targeted location in less time
10 Ibrahim, 2021 DA (Dragonfly To design a cluster formation - The proposed technique reduced The temperature of the sensor
Algorithm) technique to make efficient and network overhead and increased nodes is not considered which
minimum number clusters to cluster lifetime by forming efficient is one of the most important
make the network long-lasting and long-lasting clusters with great parameters that needs to be
energy efficiency considered in WBAN
11 Choudhary, MO-PSO To obtain minimum node MATLAB The optimal BNC location helps The issues related to cross-
Nizamuddin, (Multi- energy consumption and higher to minimize the network average channel interference in a multiple
and Zadoo, Objective network throughput during data BER and energy consumption WBAN scenario are not taken
12 Samal, Patra, ACS (Adaptive To minimize the cost and for MATLAB The proposed algorithm lowers the It results in delay as it takes
and Kabat, 2022 Cuckoo uniformly distributing load on energy consumption while taking into time for the algorithm to find an
Search)-based the relay nodes for low energy account the load on relay nodes. It optimal number of relay nodes
algorithm consumption also minimizes the cost and enhances
the network lifetime
13 Arafat, Pan, and GWO To ensure and enable energy- MATLAB The proposed model optimizes More optimal solutions can be
Bak, 2023 (Grey Wolf efficient delivery of data packets energy for inter and intra-cluster applied in the future to get more
optimization from CH to sink communication. It forms an optimal energy-efficient routing
algorithm) number of clusters and distributes
energy among the nodes significantly
which prolongs the network lifetime
234 Optimization techniques for wireless body area network routing protocols: Analysis and comparison
Thus, different optimization techniques are imple- based wireless body area networks. J. Enterp. Inform.
mented and simulated in the literature based on the Manag., 33(5), 1–22. [Link]
objective. The various parameters are defined as per 02-2020-0075.
the application requirement and for some applica- Abidi, B., Jilbab, A., and El Haziti, M. (2020). Wireless
body area networks: A comprehensive survey. J. Med.
tions multi-objective optimization model can be
Engg. Technol., 44(3), 97–107. [Link]
applied to maximum network lifetime, connectivity
0/03091902.2020.1729882.
and reliability. For cluster formation, the factors like Ali, G. A., and Al Masud, S. M. R. (2018). Routing opti-
two-hop connectivity ratio (TCR), node stability fac- mization in WBAN using bees algorithm for over-
tor (NSF) and energy factor (EF) are considered. crowded Hajj environment. Int. J. Adv. Comp. Sci.
Table 32.2 provided below helps to analyze vari- Appl., 9(5), 75–79. [Link]
ous optimization techniques used in literature to SA.2018.090510.
perform routing in WBAN in an efficient manner to Arafat, M. Y., Pan, S., and Bak, E. (2023). Distributed
solve the various network-related problems. The table energy-efficient clustering and routing for wear-
highlights the objective, optimization technique used, able IoT enabled wireless body area networks. IEEE
advantages, limitations and simulation tool used. Acc., 11, 5047–5061. [Link]
CESS.2023.3236403.
Bhatia, H., Panda, S. N., and Nagpal, D. (2020). Internet
Conclusion and future scope of things and its applications in healthcare-A sur-
vey. ICRITO 2020 - IEEE 8th Int. Conf. Reliab..
The paper examines current optimization techniques
Infocom Technol. Optim. (Trends and Future Di-
used in WBAN routing to solve many problems and rections), 305–310. [Link]
offers a systematic review of the WBAN routing TO48877.2020.9197816.
algorithms. The advantages, disadvantages, various Bhola, J., Shabaz, M., Dhiman, G., Vimal, S., Subbulakshmi,
routing parameters, optimization techniques, and P., and Soni, S. K. (2022). Performance evaluation of
implementation tools used are presented as the result multilayer clustering network using distributed energy
of the systematic review outcome. Optimization tech- efficient clustering with enhanced threshold protocol.
niques play a crucial role in improving the perfor- Wire. Pers. Comm., 126(3), 2175–2189. [Link]
mance of WBANs. These techniques aim to optimize org/10.1007/s11277-021-08780-x.
various parameters such as power consumption, data Bilandi, N., Verma, H. K., and Dhir, R. (2019). PSOBAN:
A novel particle swarm optimization based proto-
rate, network lifetime, etc. Energy efficiency, data com-
col for wireless body area networks. SN Appl. Sci.,
pression, routing protocols, scheduling algorithms,
1(11), 1–14. [Link]
and channel allocation are some of the commonly 1514-0.
used applications of the optimization techniques in Choudhary, A., Nizamuddin, M., and Zadoo, M. (2022).
WBANs. By using these techniques, the performance Body node coordinator placement algorithm for
of WBANs can be improved, and the potential of this WBAN using multi-objective swarm optimiza-
technology can be fully realized in healthcare applica- tion. IEEE Sen. J., 22(3), 2858–2867. [Link]
tions. Using an in-depth taxonomy, this article offers org/10.1109/JSEN.2021.3135269.
a better comprehension of the research concerns and Goyal, R., Mittal, N., Gupta, L., and Surana, A. (2023).
identifies advantages and key problems in the previ- Routing protocols in wireless body area net-
ous work. The aim of the survey is that it can help the works: Architecture, challenges, and classification.
Wire. Comm. Mob. Comput., 1–19. [Link]
researchers to develop new routing algorithms using
org/10.1155/2023/9229297.
optimization techniques by working on the limita-
Guleria, K., Kumar, S., and Verma, A. K. (2019). Energy
tions of the existing routing protocols in the future. aware location based routing protocols in wireless
The potential researchers will benefit as they find it sensor networks. World Sci. News, 124, 326–333.
simple to discover specific research issues and future Guleria, K. and Verma, A. K. (2019). Comprehensive review
directions from this systematic review, which will for energy efficient hierarchical routing protocols on
improve the effectiveness of routing protocols. For wireless sensor networks. Wire. Netw., 25(3), 1159–
future research; multi-objectives can be considered 1183. [Link]
and based on the objective the optimization technique Ibrahim, A. A. (2021). Quality of service-aware clustered
can be selected to achieve higher energy efficiency, triad layer architecture for critical data transmission
improved QoS, and enhanced network performance. in multi-body area network environment. Engg. Rep.,
3(7), 1–21. [Link]
Kaur, N. and Singh, S. (2017). Optimized cost effective and
References energy efficient routing protocol for wireless body
area networks. Ad Hoc Netw., 61, 65–84. [Link]
Aadil, F., Oh young Song, Mushtaq, M., Maqsood, M., org/10.1016/[Link].2017.03.008.
Sheikh, S. E., and Baber, J. (2020). An efficient cluster Kour, K. and Kang, S. S. (2019). Evaluation of wireless
optimization framework for internet of things (IoT) body area networks. Int. J. Innov. Technol. Explor.
Applied Data Science and Smart Systems 235
Engg., 8(9(Special Issue)), 350–356. [Link] Rani, S., Koundal, D., Kavita, Ijaz, M. F., Elhoseny, M., and
org/10.35940/ijitee.I1056.0789S19. Alghamdi, M. I. (2021). An optimized framework
Li, X. and Jiang, H. (2022). Energy-aware healthcare system for WSN routing in the context of industry 4.0. Sen-
for wireless body region networks in IoT environment sors (Basel, Switzerland), 21(19), 1–15. [Link]
using the whale optimization algorithm. Wire. Pers. org/10.3390/s21196474.
Comm., 126(3), 2101–2117. [Link] Saleem, F., Majeed, M. N., Iqbal, J., Waheed, J., Rauf,
s11277-021-08762-z. A., Zareei, M., and Mohamed, E. M. (2021). Ant
Mahmoud, H., Fadel, E., and Akkari, N. (2020). Routing lion optimizer based clustering algorithm for wire-
protocols in WBAN: A performance evaluation for less body area networks in livestock industry. IEEE
healthcare applications. Int. J. Adv. Res., 8(01), 334– Acc., 9, 114495–11513. [Link]
341. [Link] CESS.2021.3104643.
Mehmood, B. and Aadil, F. (2021). An efficient clustering Samal, T. K., Patra, S. C., and Kabat, M. R. (2022). An adap-
technique for wireless body area networks based on tive cuckoo search based algorithm for placement of
dragonfly optimization. Int. Things Busin. Trans. Dev. relay nodes in wireless body area networks. J. King
Engg. Busin. Strat. Indus., 5.0, 27–42. [Link] Saud University – Comp. Inform. Sci., 34(5), 1845–
org/10.1002/9781119711148.ch3. 1856. [Link]
Negra, R., Jemili, I., and Belghith, A. (2016). Wireless Seth, I., Panda, S. N., and Guleria, K. (2021). IoT based
body area networks: Applications and technologies. smart applications and recent research trends. 2021
Proc. Comp. Sci., 83(3), 1274–1281. [Link] 6th Int. Conf. Sig. Proc. Comput. Con. (ISPCC), 407–
org/10.1016/[Link].2016.04.266. 412.
Panhwar, M. A., Liang, D. Z., Memon, K. A., Khuhro, S. Shokeen, S. and Parkash, D. (2019). A systematic review of
A., Abbasi, M. A. K., Noor-ul-Ain, and Ali, Z. (2020). wireless body area network. 2019 Int. Conf. Automat.
Energy-efficient routing optimization algorithm in Comput. Technol. Manag. ICACTM 2019, 58–62.
WBANs for patient monitoring. J. Amb. Intel. Hum. [Link]
Comput., 12(7), 8069–8081. [Link] Shunmugapriya, B., Paramasivan, B., Ananthakumaran, S.,
s12652-020-02541-7. and Naskath, J. (2022). Wireless body area networks:
Qadri, Y. A., Nauman, A., Zikria, Y. B., Vasilakos, A. V., Survey of recent research trends on energy efficient
and Kim, S. W. (2020). The future of healthcare in- routing protocols and guidelines. Wire. Per. Comm.,
ternet of things: A survey of emerging technologies. 123(3), 2473–2504. [Link]
IEEE Comm. Sur. Tut., 22(2), 1121–1167. [Link] 021-09250-0.
org/10.1109/COMST.2020.2973314. Xu, G. and Wang, M. (2014). An energy-efficient routing
Rakhee, and Srinivas, M. B. (2016). Cluster based energy effi- mechanism based on genetic ant colony algorithm for
cient routing protocol using ANT colony optimization wireless body area networks. J. Netw., 9(12), 3366–
and breadth first search. Proc. Comp. Sci., 89, 124– 3372. [Link]
133. [Link]
33 Securing the boundless network: A comprehensive analysis
of threats and exploits in software defined network
Shruti Keshari1, Sunil Kumar2, Pankaj Kumar Sharma3 and
Sarvesh Tanwar4,a
Amity University Noida, Uttar Pradesh, India
1,2,4
3
ABES Engineering College, Ghaziabad, India
Abstract
Software-defined networking (SDN) is a new paradigm to increase scalability, dynamic, flexible, and programmatically
efficient configuration of networks to revolutionize network control and management via separation of the control plane
and data plane as compared to traditional networking. But this change of networking also brings some new challenges and
security issues. This research offers a comprehensive analysis of wide range of attacks faced by SDN. It starts by explaining
the detailed architecture of SDN along with their work flow which leads to the possible security challenges and threats. It
also provides detailed analysis of numerous attacks, such as data plane attack, controller centric assaults and possible vulner-
abilities in each plane that helps attackers to inject malware or exploit the weaknesses in SDN. Every threat is scrutinized in
depth by defining each attack methods with their tools. It further discusses how SDN security risks are changing, taking into
account possible new risks and developments. To sum up, this study provides an invaluable tool for researchers, practitioners,
and network security experts who want to comprehend potential weaknesses and their implications at various degrees. It
seeks to support ongoing efforts to strengthen the security of SDN infrastructures in an ever-evolving cybersecurity ecosys-
tem by thoroughly examining the threat landscape and their impact.
Keywords: Software defined network, security challenges, attacks, tools and techniques used by attacker, vulnerabilities
s.tanwar1521@[Link]
a
Applied Data Science and Smart Systems 237
software programmability also benefited networking is giving a detailed view of SDN architecture (Singh
by making implementation of complex system with et al., 2019; Jimenez et al., 2021), with its services.
simple software routine and algorithm. Figure 33.2 Figure 33.3 is defining work flow of SDN, as how
238 Securing the boundless network: A comprehensive analysis of threats and exploits in software
Data plane
In SDN architecture, the data plane is responsible for
data packet forwarding in accordance with instruc-
tions from the SDN controller. The forwarding fea-
ture is implemented by data plane devices, which are
usually switches or routers, and flow tables are kept
up to date to control how traffic is handled. For the
purpose of receiving flow controls and reporting net-
work information, they speak with the SDN control-
ler by using northbound API.
Southbound interface
The communication channel between the data plane
devices and the SDN controller is known as the south-
Figure 33.3 Work flow of SDN bound interface. It gives the controller the ability to
communicate with network devices, set up flow rules,
and gather network state data. OpenFlow is the most
decoupling of control plane and data plane is mak- popular southbound protocol used in SDN, however,
ing network more efficient. A detailed discussion of P4 (Liatifis et al., 2023) and NETCONF (Kunz et al.,
involved interfaces with aforementioned three layers 2017) are also employed.
is provided in this section.
Northbound interface
Application plane The northbound interface allows for communica-
Application plane is also known as application layer. tion between higher-level network applications
Application plane is responsible for utilizing the capa- or orchestration systems and the SDN control-
bilities provided by the SDN controller to implement ler. Applications can use it to set network policies,
specific network services or policies. These applica- ask the controller for network state information,
tions can be developed by network administrators, and request network services. The northbound
third-party developers, or vendors. They leverage interface allows application developers to connect
the programmability of the SDN infrastructure to programmatically with the network infrastruc-
dynamically configure network behavior, optimize ture by abstracting away the underlying network
traffic flows, or implement network security mea- complexity.
sures. Application layer also handle orchestration and
service chaining with service innovation.
Security challenges in SDN
Control plane Analysis of security attacks will become easier if
Control plane is responsible for management and the objective of that attack is clear. Intention of the
control of whole network. The core element of SDN attacker is key point for detecting and preventing
architecture is the SDN controller, which resides in network system from these attacks. Just like flooding
this plane. Controller is the heart of SDN architec- of false messages, shows the intention of attacker to
ture. Controller is in charge managing and controlling affect the performance of system and also the avail-
network devices like switches, router and firewall. ability of the resources. Understanding the security
Controller gives a centralized view of the network issues makes it easier to spot and address these prob-
and enforces network policies. The controller also lems. This section is describing the issues and threats
configures flow rules and controls network traffic by of network security in SDN context (Chica et al.,
interacting with network devices using protocols like 2020).
OpenFlow. Control plane is responsible for managing Data privacy and confidentiality: Massive amounts
network state database, network application and con- of sensitive data are handled in SDN setups. This
trol protocol. Control plane plays an important role includes user data, financial information, medical
for communication between two other two planes. Its information, and intellectual property. It is essential
Applied Data Science and Smart Systems 239
impact on network, where as some attackers use pre- classified based on different factors, but it is not an
defined techniques which are also useful in other net- exhaustive list. The evolving nature of SDN technol-
work also. ogy may introduce new attack vectors and techniques
It’s important to note that attacks can often fall that may require further categorization and analy-
into multiple categories, and the classification of sis. There are some attacks which are affecting mul-
attacks may vary depending on the specific context tiple planes (Abd Elazim et al., 2018; Hegazy et al.
and perspective of the analysis. This categorization 2021; Alhaj et al., 2022) simultaneously. Table 33.4
provides a high-level overview of how attacks can be is giving the view of these types of attacks with their
242 Securing the boundless network: A comprehensive analysis of threats and exploits in software
attacked area, plane and affected functionality. For Hegazy et al., 2021). Knowledge of these vectors and
better understanding of these attacks, Table 33.4 is techniques helps to improve security of the system.
also giving brief view about these attacks with one Table 33.5 is giving summary about these attack
example. vectors with their method and technique. For better
These examples illustrate how each attack type can understanding of this, Table 33.5 is also giving brief
be carried out in an SDN environment. Attackers can about their effect and latest tools used by the attacker.
employ variations and combinations of these attacks, Disease may be transferred from patients to doctors
and the specific techniques used may vary depending and vice-a-versa.
on the attacker’s goals and the vulnerabilities present
in the SDN infrastructure. Vulnerabilities and exploitable weaknesses
Analysis of attacks and their various attacking
Attack vectors and techniques
tool, conclude that SDN network is not immune
Attack vectors and techniques is a pathway or method to the vulnerabilities. Strong points of SDN i.e.,
used by the attackers for illegal access of network programmed network devices, central control
and launch attacks in SDN (Mahajan et al., 2020; point and dynamic adaptability of this network
is opening new weak points for attackers. Attacks Table 33.6 SDN component with their threats
defined in Tables 33.4 and 33.5 leads the way to
Component Threats
get the knowledge of weaknesses and areas, which
needs to be strong for making system more secure. SDN controllers • Insecure authentication
Knowledge of these vulnerabilities will help to • Vulnerable software
understand intension of attacker. Some common • Lack of secure communication
exploit weaknesses are: Network devices • Firmware vulnerabilities
(Switches, Routers, • Weak access controls
• Insecure controller communication etc.) • Lack of flow rule validation
• Controller software vulnerabilities
Communication • Lack of encryption
• Weak access controls channels • Inadequate authentication and
• Insecure southbound interfaces authorization
• Flow rule manipulation • Protocol-level vulnerabilities
• Insufficient monitoring and logging
• Lack of network segmentation.
Vulnerabilities may arise due to many factors of • Software define security services
the system. Some weaknesses come due to their archi- • Authentication and access control
tectural characteristics of SDN. Table 33.6 is defin- • Encryption and privacy preservation
ing such threats which arise due to architectural • Security orchestration and automation
components. • Collaboration and threat intelligence sharing
Abstract
Individuals with communication disorders has dramatically increased over the past two decades. Similarly due to the qua-
drupling in publications since 2012, a bibliometric study is needed for an hour. This study offers a thorough analysis of com-
munication disorders across an appropriate period. We have utilized bibliometrics to assess communication disorder-related
articles published in the Scopus database between 1960 and 2022, and to illustrate the resulting rise in research publications
based on a number of factors, including (i) publishing patterns (e.g., contributing authors, affiliations), (ii) key term analysis
to identify domain of interest, (iii) key term bunching, (iv) citation patterns, (v) publications medium, and (vi) researchers
who assist in examining research productivity in this particular domain. Based on the Scopus database, a total of 80,289
papers about communication disorders were examined. In the end, 59,252 publications and 12,232 key terms were retained,
especially those related to communication disorders. The number of publications increased by 77.7% (60 in 1960, 4725 in
2022). The United States contributes the most publications, and the highest document type is articles. Medicine (53,294)
has the most documents, followed by psychology (15,928) and so forth engineering (2287). Over time, the relative weights
of various study fields have also altered. A meta-perspective literature review is conducted on the quantitative characteristics
and properties of communication disorders. The suggested analytical study will be a vital resource for a substantive discus-
sion about potential future research plans for supporting special people with communication disorders.
muskaanchawla07@[Link]
a
Applied Data Science and Smart Systems 247
examining the distribution of data based on the fre- Bibliometric techniques are pivotal for producing
quency count and takes note of the citation’s style. innovative insights, such as assessing the productiv-
Moreover, it considers the quality, structure, and ity of writers, algorithms, or key term clusters (Calvo,
exchange of information in literature affiliation and Carbonell, and Johnsen, 2019; Zhang et al., 2022). In
the research’s financial impact (Hood and Wilson, order to give unique insights, the objective is to char-
2001). acterize the current state of communication disorder
As a result, the survey of literature revealed that in a pertinent time period. A range of techniques are
less research has been done in this field. Based on the being used, including computational and quantitative
Scopus database, the productivity of authors and con- algorithms, to analyze important factors (Hood and
tributing nations have been examined for 35,120 arti- Wilson, 2001). As shown in Figure 35.2, the num-
cles relevant to communication disorders during the ber of publications on communication disorders has
years of 2011 and 2022. Key phrase clusters associ- been exponentially rising every year since 2002. Every
ated with subject areas and various opinions of publi- year since 2002, a search result has increased by two,
cation trends, research impact, and productivity have according to Google Scholar (2022). The same rise is
been examined by the authors Heilig and Vob (2014). also being anticipated by a site for scientific literature
However, it falls short in terms of giving information (Scopus). Scopus counts 80,289 pertinent publications
on communication disorder research trends, citation as of 12 September 2022, whereas communication
patterns, and most-cited publications. Since most of disorder includes 59,252 disorders related papers. To
the analysis is based on a low number of publica- the best of our knowledge, no research has been done
tions for a certain field, and straight count technique in the area of bibliometric assessments in this field.
(Zhang et al., 2022). Based on the Scopus database, 80,289 papers relat-
ing to communication impairments are being examined
in this article (1960–2022). This study counts peer-
reviewed papers empirically and statistically. Current
research is still employed as a broad experimental
premise, which may stimulate academics to conduct in-
depth bibliometric investigations in various scholastic
orders. In order to identify a subject of interest, this
work aims to provide insight on (i) publishing patterns
(such as contributing authors and affiliations), (ii) an
analysis of popular key terms, and (iii) key term bunch-
ing. Other information, such as (i) journal citation pat-
terns, (ii) sources of publications, (iii) affiliations, and
Figure 34.1 Different communication disorders with (iv) the authors’ insights, aids in examining the research
their prevalence
output of affiliations and researchers. As a result, the Data is being initially gathered through Scopus
knowledge in this work is being provided from the var- databsae and pre-processed before presenting pat-
ious aspects of evolution, status, and trends. tern analysis (removal of irrevalant and duplicate
The remaining document is structured as follows. documents). After the screening stage, the abstract of
In methodology section, dataset pre-processing, col- various articles have been reviewed and excluded the
lection of dataset, research trend, productivity and records that were not related to the field. During the
contribution to research are briefly described. Based eligibility stage, various full text articles have been
on important phrase clusters, research analyses sec- evaluated and removal of thesis or arVix publications
tion observes the keyword analyses and cluster for- have been done. This section serves as an example of
mation related to communication disorder followed the elaborated strategy used to locate peer-reviewed
with analyses on the basis of authors name and affilia- articles published over the last 10 years as shown in
tion section. Analysis is done on the basis of scholarly Figure 35.3.
metrics such as CiteScore per year and source normal-
ized impact per paper (SNIP) has been analyzed fol- Bibliographic analyses plan
lowed by challenges and the conclusion of the work. In order to obtain and process organized information
of articles, Elsevier’s Scopus database has been used
Methodology as a manual treatment of bibliographic information
(Serenko and Bontis, 2004; Garg, Sidhu, and Rani,
Owing to the specific field (communication disorders) 2019). In comparison to other bibliographic data-
under investigation, the analysis is based on 80,289 bases, Scopus provides advanced functionality for
articles, which points in the direction of limited infer- exporting organized information, including biblio-
ence at this stage. The analysis contains extensive time graphic and reference data, abstracts, and key terms.
of observation and employs a range of techniques to Scopus has twice as many articles for this domain as
analyze important factors. This section serves as an WoS, which incorporates articles-in-press for publi-
example of the elaborate strategy used to locate peer- cations (Zhang et al., 2022). In the section, research
reviewed articles published over a span of various trend, contributions and productivity of the research
years. have been discussed.
abstract, and key terms to encompass extensive lit- clusters in order to examine important subjects and
erature in communication disorder. A list of 80,289 aspects. The substance of literature that is connected
items has been produced as a consequence over the to a given domain is categorized using key phrases.
observation period of 1912–2022 (as of 24 September Major words define key themes and distinctive
2022). In order to find articles with correct and com- research contribution qualities. The frequency and
prehensive data and metadata (authors, titles, double pace of important phrases within a given time period
entries, etc., are deleted), a data cleansing process is can be used to identify the rise of current study sub-
carried out. Last but not least, 59,252 publications jects. It is also possible to identify themes or points
and 12,232 key phrases that were directly relevant to of view that are closely related to one another by
communication disorder has been kept; 85.64% of looking at the frequency of important phrases. The
publications were written in English. Data has been analysis is carried out in the following subsections,
divided into segments and examined from various combined the terms depending on how frequently
angles. First, the publications are being examined to they occur together.
determine which disciplines have contributed to the
development of the topic. Second, the number of pub- Keyword analyses
lications and the total number of citations are used This review has been conducted by using Scopus data
to analyze the contributing nations. The analysis also sources. The data was gathered and processed from
covers the distribution of document types, the number organized information of articles. The growth rate of
of publications per outlet, and the number of authors annual scientific production is 20.04%. The cumula-
who contributed to each article. tive number of articles containing the keyword by
year is shown in Figure 35.4. The cumulative occur-
Research contribution rence of related keywords like autism, children, com-
Researchers from numerous disciplines are exerting munication, disorders, etc., is showing a significant
significant effort to advance knowledge. Research increase per year. This role in the performance that
publications are a primary source of information for for 25.33% of the publications under study, key
topics that are the subject of scientific inquiry. Every terms have been assigned by Scopus, whereas key
research article has several objectives that define its terms are defined for 74.67% of the publications
contribution to the field. under study.
There are numerous distinct research contribu- Regarding the typical distribution of key phrases
tions. However, there are numerous other types of per publication, it has been found that the list of the
contributions that result in new truths. On the basis publication’s objectives typically uses three to six
of (i) academic contribution, (ii) contributing nation, key terms. However, when there is a communication
(iii) number of researchers, (iv) average number of issue, the frequency and number of crucial terms are
citations, and (v) research outlet, the examples in the higher than what is often believed. The ultimate goal
subsequent sections can be classified into three dis- is to reduce key phrase inconsistency; publishers fre-
tinct categories of contributions (conference, book quently supply a list of key terms pertinent for a par-
chapter, and journal). ticular journal or conference.
Analyses on the basis of authors names and academic fields, as well as the areas, authors, citation
affiliation patterns, and sources that contribute to it. Additionally,
subsections following indicates (i) the average num-
This section examines the general layout and advance- ber of articles by the eminent researchers, (ii) explores
ment of communication disorders across a range of publications by country and authorship patterns to get
Applied Data Science and Smart Systems 251
insights from contribution patterns and (iii) examines in, using data derived from Scopus. In addition to
how publications’ impact is influenced by the outlet. the fact that contributions to computer science are
ongoing, there is also evidence of a popularity peak
Publications by various authors and a lack of expectations in the field of psychiatry.
Figure 35.6 displays the average number of docu- It suggests that ongoing expectations have an impact
ments by the top 10 authors in research of commu- on current research (Li et al., 2021). Data in Figure
nication disorders (across a decade). Compared to 35.7 shows that medicine and psychiatry have been
other authors’ contributions, a contribution by Lord, made significant contributions. Thus, it demonstrates
C. is noteworthy (128). In contrast to the individual, that study is concerned with having a thorough
it appears that communication dysfunction is more understanding of a particular field. It offers benefits
prevalent in collaborative settings. Over research in the domains of computer science and engineering
conducted by a single researcher, a certain group of in addition to the trends of communication disorders
authors may have a legitimacy advantage. (García-Aroca et al., 2017). This section investigates
publications by country and authorship patterns to
Evaluation based on academic disciplines and con- gain insights from contribution patterns. According
tributing country to Figure 35.8, researchers from the United States
Each article has been divided into multiple catego- (38.10%) and the United Kingdom (11.23%) con-
ries according to the publication outlets it appeared ducted the majority of the research. It is clear that
Table 34.1 Metadata about the research contributions Articles using scholarly metrics
Document type Overall (in percentage) This section analyses the scientific output and influence
of authors, institutions, and nations by Scopus. Scopus
Conference paper 43.49 can be used to gauge a journal’s prominence inside the
Article 72.6 database. The following metrics are used by Scopus
Book chapter 18.03 journal analyzer to evaluate journals and articles.
Review 14.86
CiteScore per year
Editorial 16.3
CiteScore is obtained by dividing the number of
Short Survey 0.819 Scopus-indexed papers published in the same three
Book 0.313 years by the number of citations that publica-
tion received in a given year. Table 34.2 depicts the
CiteScore per year of renowned journals during the
time period 2011–2021.
the United States (38.10%) and the United Kingdom
(11.23%) make considerable contributions to this Source normalized impact per paper (SNIP)
domain. SNIP is a correction metric that takes into consid-
eration the variations in citation potential between
Publication outlet fields. Table 34.3 depicts the SNIP of the eminent
The visibility and impact of an item is being influ- journals for the duration of last 10 years.
enced by the outlet choice. This section examines
how the research community’s outlets like to share SCImago journal rank
their ideas and expertise. It is feasible to analyze It indicates the typical number of weighted citations
the data since it contains metadata about the type that the papers published in the chosen journal over
of document. According to Table 34.1, 43.49% the preceding 3 years obtained in the chosen year.
of articles are, on average, included in conference Table 34.4 depicts the SJR of the esteemed journals
proceedings. during the time period of last 10 years.
Table 34.1 makes clear the facts that, the major-
ity of research contributions are presented and pub- Challenges and opportunities issues
lished in journals (72.76%), as opposed to books
(18.03%) and editorial (16.3%). The impact of The present state of research cannot easily be inferred
citations on the volume of research and publica- from surveys or literature reviews. As a result, the
tions per journal is discussed in the following lead- methodology or strategy utilized in this analysis can
ing section. be applied to any field of study. The relationship
Applied Data Science and Smart Systems 253
Table 34.2 CiteScore per year
Source 2011 2012 2013 2014 2015 2016 2017 2018 2019 2020 2021
Journal of Autism and 6.2 5.8 5.8 5.9 6.6 6.5 6.2 5.6 5.2 5.6 6.6
Developmental Disorders
Autism 4.8 4.2 4.2 4.4 5.3 6.7 7.7 7.8 6.8 6.7 7.5
PLoS One 4.5 4.1 4.4 5.1 5.6 5.9 5.7 5.4 5.2 5.3 5.6
International Journal 2.8 3.1 2.8 2.8 3.3 3.9 4 2.8 3 3.5 4.4
of Language and
Communication Disorders
Research in Autism Spectrum 3.2 4.2 4.2 4.8 4.4 3.8 3.7 3.1 3.1 2.8 3.7
Disorders
Source 2011 2012 2013 2014 2015 2016 2017 2018 2019 2020 2021
Journal of Autism 1.752 1.884 1.72 1.635 1.585 1.565 1.52 1.564 1.53 1.473 1.862
and Developmental
Disorders
Autism 1.385 1.454 1.237 1.514 1.369 1.557 1.728 1.872 2.14 1.93 2.15
PLoS One 1.256 1.175 1.168 1.144 1.158 1.124 1.153 1.179 1.197 1.322 1.368
International Journal 1.515 1.231 1.222 1.211 1.414 1.541 1.48 1.268 1.082 1.358 1.845
of Language and
Communication
Disorders
Research in Autism 1.291 1.098 1.094 1.143 0.952 0.817 0.84 0.861 1.017 1.085 1.243
Spectrum Disorders
Source 2011 2012 2013 2014 2015 2016 2017 2018 2019 2020 2021
Journal of Autism and 1.835 1.821 1.805 1.947 1.976 1.955 1.81 1.675 1.434 1.374 1.207
Developmental Disorders
Autism 1.228 0.993 1.129 1.503 1.455 1.844 1.739 2.336 1.885 1.899 1.617
PLoS ONE 2.425 1.982 1.772 1.559 1.427 1.236 1.164 1.1 1.023 0.99 0.852
International Journal 1.022 0.999 0.809 0.796 1.049 1.222 1.057 0.807 0.821 1.101 0.95
of Language and
Communication Disorders
Research in Autism 0.824 1.032 0.989 1.368 1.063 0.86 0.844 0.872 0.834 1.04 0.895
Spectrum Disorders
between authors and topics might be analyzed as technology as follows: (i) Clinical assessments and
part of future work to identify patterns in network proctored exams for the current diagnosis are subjec-
structure or trends within the domain. In order to tive and less technologically supported (Taylor and
produce new outcomes or assessments, the findings Whitehouse, 2016; Mandy et al., 2017; Ellis Weismer
or outcomes of this article can be compared to other et al., 2021). As a result, technical diagnostic instru-
results. It is possible to enhance the methods used ments are required. (ii) Additionally, conventional
for pre-processing research data so that less manual intervention approaches depend on the therapists’
work is required. Apart from these challenges, there knowledge. Due to waiting periods and costs, the
are few more opportunities issues related to the procedures are increasingly routine; as a result, early
254 A bibliometric analyses on emerging trends in communication disorder
Table 34.5 Meta-perspectives of the study technique. Int. J. Dev. Neurosci., 26(7), 699–704.
[Link]
S. No. Parameter Frequency Barnard-Brak, L., Richman, D. M., Chesnut, S. R., and Lit-
tle, T. D. (2016). Social communication questionnaire
1 Maximum people suffering Speech problems scoring procedures for autism spectrum disorder and
from the prevalence of potential social communication dis-
2 Significant rise in publication 2002 order in ASD. School Psychol. Quart., 31(4), 522–533.
from [Link]
3. Highest publications United States Mahajan, Anmol, and Abhinav Bhandari. (2012). Attacks
contributed by country in Software-Defined Networking: A Review. In Pro-
4. Highest document by type Articles (72.6%) ceedings of the International Conference on Innova-
tive Computing & Communications (ICICC. 1–10.
5. Average CiteScore 1.33
Braithwaite Stuart, L., Jones, C. H., and Windle, G. (2022).
6. Average SNIP 1.386 A qualitative systematic review of the role of families
7. Average SJR 1.33 in supporting communication in people with demen-
tia. Int. J. Lang. Comm. Dis., 1130–53. [Link]
org/10.1111/1460-6984.12738.
Calvo, F., Carbonell, X., and Johnsen, S. (2019). Information
assessment is not possible, which delays interven- and communication technologies, e-health and home-
tion times and lowers the quality of life for both lessness: A bibliometric review. Cogent Psychol., 6(1).
children and parents (Arthi and Tamilarasi, 2008; [Link]
Barnard-Brak et al., 2016; Taylor and Whitehouse, Centre for disease control and prevention. (2022). https://
2016). Caretakers may do biased evaluation despite [Link]/nchs/products/databriefs/[Link].
Ellis Weismer, S., Rubenstein, E., Wiggins, L., and Durkin,
the availability of subjective rating scales due to igno-
M. S. (2021). A preliminary epidemiologic study of
rance. As a result, diagnostics should be automated.
social (pragmatic) communication disorder relative
(iii) There is a need to assist in the creation of edu- to autism spectrum disorder and developmental dis-
cational supports and intervention programs because ability without social communication deficits. J. Aut.
there are not enough patients receiving proper and Dev. Dis., 51(8), 2686–2696. [Link]
multimodal treatment (Gonzalo et al., 2019; Herrero s10803-020-04737-4.
and Lorenzo, 2020). García-Aroca, M. Á., Pandiella-Dominique, A., Navar-
ro-Suay, R., Alonso-Arroyo, A., Granda-Orive, J.
I., Anguita-Rodríguez, F., and López-García, A.
Conclusion (2017). Analysis of production, impact, and scien-
This study examined the literature using meta-per- tific collaboration on difficult airway through the
spectives to analyze the quantitative dimensions and web of science and Scopus (1981–2013). Anesth.
traits of communication disorders. The article gives Anal., 124(6), 1886–1896. [Link]
ANE.0000000000002058.
a thorough overview of communication impairments
Garg, D., Sidhu, J., and Rani, S. (2019). Emerging trends in
over the appropriate time period. Based on the Scopus
cloud computing security: A bibliometric analyses. IET
database, a total of 59,252 publications concerning Software, 13(3), 223–231. [Link]
the communication disorders, a number of scientific sen.2018.5222.
publications and significant contributions science Gonzalo, L., Lledó, A., Arráez-Vera, G., Lorenzo-Lledó, A.
related to computer science and psychiatry are exam- (2019). The application of immersive virtual reality
ined. Table 34.5 concludes all the meta-perspectives for students with ASD: A review between 1990–2017.
of the study. Educ. Inf. Technol., 24, 10639.
Google Scholar. (2022). [Link]
scholar?hl=en&as_sdt=0%2C5&q=communication+
References disorder+bibliographic+review&oq=communication+
disord.
Adams, C., Lockton, E., Freed, J., Gaile, J., Earl, G., Mc- Heilig, L. and Vob, S. (2014). A scientometric analysis
Bean, K., Nash, M., Green, J., Vail, A., and Law, J. of cloud computing literature. IEEE Trans. Cloud
(2012). The social communication intervention proj- Comput., 2, 266–278. [Link]
ect: A randomized controlled trial of the effectiveness TCC.2014.2321168.
of speech and language therapy for school-age chil- Herrero, J. F. and Lorenzo, G. (2020). An immersive virtual
dren who have pragmatic and social communication reality educational intervention on people with autism
problems with or without autism spectrum disorder. spectrum disorders (ASD) for the development of com-
Int. J. Lang. Comm. Dis., 47(3), 233–244. [Link] munication skills and problem solving. Educ. Inform.
org/10.1111/j.1460-6984.2011.00146.x. Technol., 25(3), 1689–1722. [Link]
Arthi, K. and Tamilarasi, A. (2008). Prediction of autistic s10639-019-10050-0.
disorder using neuro fuzzy system by applying ANN
Applied Data Science and Smart Systems 255
Hood, W. W. and Wilson, C. S. (2001). The literature social robotics: A systematic review. Aut. Res., 9(2),
of bibliometrics, scientometrics, and informet- 9–11.
rics. Scientometrics, 52(2), 291–314. [Link] Van Raan, A. (1997). Scientometrics : State-of-the-art. Sci-
org/10.1023/A:1017919924342. entometrics, 38(1), 205–218.
Howard, G. S., Cole, D. A., and Maxwell, S. E. (1987). Re- Rajendran, P., Jeyshankar, R., and Elango, B. (2011). Sci-
search productivity in psychology based on publica- entometric analysis of contributions to journal of sci-
tion in the journals of the American psychological as- entific and industrial research. Int. J. Dig. Lib. Ser.,
sociation. Am. Psychol., 42(11), 975–986. [Link] 79–89.
org/10.1037/0003-066X.42.11.975. Schwarze, S., Voß, S., Zhou, G., and Zhou, G. (2012). Sci-
Jensen de López, Kristine, M., Kraljević, J. K., and Struntze, entometric analysis of container terminals and ports
E. L. B. (2022). Efficacy, model of delivery, intensity literature and interaction with publications on distri-
and targets of pragmatic interventions for children bution networks. Lec. Notes Comp. Sci., 7555, 33–52.
with developmental language disorder: A systematic [Link]
review. Int. J. Lang. Comm. Dis., 57(4), 764–781. Serenko, Alexander, and Nick Bontis. (2004). Meta-review
[Link] of knowledge management and intellectual capital
Lewis, B. R., Templeton, G. F., and Luo, X. (2007). A scien- literature: Citation impact and research productivity
tometric investigation into the validity of IS journal rankings. Knowledge and process management, 11(3),
quality measures. J. Assoc. Inform. Sys., 8(12), 619– 185–198. [Link]
633. [Link] Taylor, L. J. and Whitehouse, A. J. O. (2016). Autism spec-
Li, W. S., Yan, Q., Chen, W. T., Li, G. Y., and Cong, L. trum disorder, language disorder, and social (prag-
(2021). Global research trends in robotic applications matic) communication disorder : Overlaps, distin-
in spinal medicine: A systematic bibliometric analy- guishing features, and clinical implications. Aus.
sis. World Neurosurg., 155, e778–e785. [Link] Psychol., 51(4), 287–295. [Link]
org/10.1016/[Link].2021.08.139. ap.12222.
F. P. Mahabalagiri N. Hegde (2021). Assessment of commu- T. R., R., Lilhore, U. K., M, P., Simaiya, Kaur, and Hamdi,
nication disorders in adults. 4th ed. Plural Publishing, M. (2022). Predictive analysis of heart diseases with
Incorporated, 1–444. machine learning approaches. Malaysian J. Comp.
Mandy, W., Wang, A., Lee, I., and Skuse, D. (2017). Evalu- Sci., 1, 132–148.
ating social (pragmatic) communication disorder. J. Zhang, S., Wang, S., Liu, R., Dong, H., Zhang, X., and Tai,
Child Psychol. Psychiat. Allied Dis., 58(10), 1166– X. (2022). A bibliometric analysis of research trends
1175. [Link] of artificial intelligence in the treatment of autistic
Pennisi, P., Tonacci, A., Tartarisco, G., Billeci, L., Ruta, spectrum disorders. Fron. Psychiat., 13. [Link]
L., Gangemi, S., and Pioggia, G. (2016). Autism and org/10.3389/fpsyt.2022.967074.
35 Enhancing latency performance in fog computing through
intelligent resource allocation and Cuckoo search
optimization
Meena Rani, Kalpna Guleriaa and Surya Narayan Panda
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
Applications for the internet of things (IoT) have rapidly expanded, posing a number of difficulties in terms of quality of
service (QoS), delay, latency, and disconnections. These difficulties are still there even if fog computing has emerged. An in-
novative resource allocation strategy is presented for fog computing’s latency problems. The proposed approach includes
allocating requests and assigning jobs to improve technical capabilities. It uses a queuing system with 3 recommended lists
(a) block, (b)wait, and (c) scheduling – that is on the basis of loaded criteria to choose available nodes. The Cuckoo search
(CS) approach, which significantly lowers latency, is developed to enhance resource allocation by calculating the distance
between fog nodes and users and selecting the nearest along with the available nodes for request processing. By contrasting
latency measures with and without the CS algorithm, the given evaluation shows the value of strategy. The results show a
striking drop in latency, with the method having the distance and decreasing overall latency. The cuckoo search method is
integrated with the suggested resource allocation mechanism to produce measurable latency reductions and improves the
general performance of fog computing systems to enhance quality of life.
Keywords: Resource allocation, fog computing, internet of things, Cuckoo search algorithm, resource use efficiency
[Link]@[Link]
a
Applied Data Science and Smart Systems 257
For optimum performance and to improve the QoS resource management, decreased reaction time, and
and QoE for end users, these issues must be resolved was more effective at scheduling and controlling fog
(Moshref et al., 2022). Queuing scheduling, as well as devices (Vashisht et al., 2022). Another study put out
algorithms of load-balancing, along with the adop- a strategy for managing resources in vehicle fog com-
tion of energy-efficient procedures, may be used as puting that relies on pricing-based stable algorithms
solutions to these problems (Hong et al., 2018). to match contracts and encourage the sharing of
To reduce latency, the study suggests a brand-new resources across cars. This approach, which is compa-
queue list technique for scheduling a task inside the rable to existing optimum search algorithms but less
base broker in fog computing settings. The strategy sophisticated, demonstrated good resource manage-
makes use of 3 lists depending on a loaded set of vari- ment outcomes (Zhou et al., 2019). Researchers have
ables to estimate node availability. The CS method is suggested a method for allocating resources based on
utilized to minimize latency and optimize resource the development of an effective work scheduling algo-
allocation. CS assesses solutions according to fit- rithm. The technique lowered energy use, increased
ness, replaces the less fit ones, and preserves popu- bandwidth use, and improved reaction times for
lation variety. The technique assists in distributing internet-dependent applications (Jamil et al., 2020).
computational resources across fog nodes by taking An enhanced type of the fundamental ant colony
into account things like loaded factors along with optimization method for work scheduling was car-
shortest pathways. The following primary contribu- ried out by researchers in 2020. The investigators
tions are presented in this study to improve resource compared the new ant colony optimization method
management. against the original ACO algorithm using a MATLAB
It suggests a method of allocating requests along application. The outcomes demonstrated that the
with scheduling work on the basis of grouping them upgraded method increased the overall completion
into 3 categories (wait, block, scheduling) according time and economic cost, consequently enhancing the
to the amount of load. The article offers allocation of service quality in a fog setting (Yin et al., 2020). In
a resource method which reports the latency issues in 2021, investigators put out the TRAM method, which
fog computing by using the CS methods to find the uses the expectation-maximization method to sched-
optimal allocation of resources. ule activities in a fog setting. TRAM is a new method
The CS method is used in this study to discover the for the allocation of a resource along with manage-
closest fog nodes with the shortest pathways while tak- ment within fog computing. The major objectives of
ing into account node availability and user proximity resource allocation in a fog setting are to optimize
to choose the best nodes to handle queries. The objec- energy usage and job distribution. The scholars uti-
tives of these contributions are to reduce latency and lized iFogSim to test the performance and discovered
boost fog computing system performance. Following that TRAM led to a 60% enhancement in time of exe-
is an organization of the remaining article. An over- cution for assigning resources (Wadhwa et al., 2022).
view of fog computing, followed by present methods Another study recommended utilizing a synthetic
for optimizing resource allotment in fog computing, ecosystem-based optimization technique in com-
and concludes with finally which contains test results. bination with the SSA to optimize the process of a
task-scheduling. The suggested method, known as
AEOSSA, performed better in terms of productivity
Related work
and time rate than previous metaheuristic approaches
Investigators presented a method for task scheduling (Abd Elaziz et al., 2021). Finally, to minimize latency
on the basis of container characteristics in smart pro- and boost service quality in the fog environment,
duction in 2018. The method demonstrated a 10% researchers suggested a queuing theory-based CS
reduction in execution time and a 5% increase in approach in 2022. They employed particular meth-
task capacity in the fog environment (Member et al., ods to gauge the level of service and optimized this
2018). Researchers put out allocation of a resource method with the CS algorithm. The findings revealed
method which adopted security and privacy con- that utilizing the CloudSim simulator enhanced aver-
cerns in fog computing the same year. The suggested age reaction time by 20.39% and reduced energy
method enhanced the robustness and security of fog usage by 12.55% (Iyapparaja et al., 2022) shown in
nodes (Zhang et al., 2018). Table 35.1.
A group of academics suggested a bio-inspired Finally, discovered that the cuckoo algorithm is
hybrid method for work planning and resource man- superior as compared to methods like the genetic algo-
agement within fog devices. It combines modified rithm in terms of selecting further and more solutions.
particle swarm optimization (MPSO) and modified This was after performing an integrated study on
cat swarm optimization (MCSO) techniques. In com- every technique and approach utilized to enhance the
parison to existing algorithms, the method enhanced allocation of a resource inside the fog environment.
258 Enhancing latency performance in fog computing through intelligent resource allocation
Table 35.1 Literature survey summary
(Zhang et al., Genetic algorithm iFogSim and fog Enhancing QoS- Increasing latency, bandwidth,
2023) environment aware scheduling utilization, and reducing
energy consumption
(Saif et al., Grey Wolf optimizer CloudSim and cloud- Improve task Reduces energy consumption
2023) fog environment scheduling and delay
(Singh and Bio-inspired hybrid MPSO and MCSO Manage resources Reduced response time,
Singh, 2022) algorithm and schedule tasks enhanced resource
inside fog devices management, and greater
efficacy by comparing it to
other algorithms
(Iyapparaja et QTCS (queuing CS algorithm Decrease latency An increase in average
al., 2022) theory-based CS) and increase service response time of 20.39% and
quality a reduction in energy use of
12.55%
(Wadhwa and Expectation EM algorithm m Schedule tasks in a 60% reduction in the amount
Aron, 2022) maximization (EM) fog setting of time required to allocate
algorithm resources; reductions in both
energy usage and task load;
and improvements in task
allocation
(Abd Elaziz, AEO (artificial SSA (salp swarm Optimize the task Performed other metaheuristic
Abualigah, and ecosystem-based algorithm) scheduling process techniques in terms of
Attiya, 2021) optimization) productivity and rate of time
(Jamil et al., Efficient job Shortest job first Increase the response Compared to other algorithms,
2020) scheduling algorithm (SJF) time of internet- the average response time was
reliant apps while enhanced by 32%, and the
simultaneously energy usage was enhanced by
lowering energy 16%
usage and improving
bandwidth utilization
(Yin et al., Improved ant colony ACO Improve the The total and economic cost
2020) optimization (ACO) completion of the as well as the amount of time
algorithm task in terms of the needed to finish have been
overall cost reduced, which has led to an
increase in the overall quality
of the service
(Zhou et al., Pricing-based Stable algorithms on Efficient resource Good findings in resource
2019) resource management the basis of pricing management management and the same
for the vehicle fog in another optimal search
computing methods but less complex
(Member and Container-based N/A Reduce execution The capacity of the task
Luo, 2018) scheduling time and improve increased by 5%, while the
task distribution time of execution was reduced
by 10%
(Zhang and Li, Privacy-focused N/A Increase the fog Guaranteeing user privacy as
2018) resource allocation nodes’ robustness well as resolving security issues
and security
The CS has a larger search space than the other algo- System model and methodology
rithms because of its solutions (called nests), which
Three system models called fog computing, resource
resemble the fog node distribution. As a consequence,
allocation and cuckoo search algorithms with its
the alternatives are more complicated and optimal.
methodologies are introduced here which are given
Now chose to base the distance in the proposal on
below:
its ability to improve the task, and this decision pro-
duced superior outcomes from other efforts.
Applied Data Science and Smart Systems 259
Optimizing resource allocation in fog computing tasks and distributes them to the relevant nodes in
uses the proposed mechanism accordance with the suggested queuing mechanism
that is made up of three lists, as previously men-
The cuckoo algorithm is suggested as a way to tioned. All forthcoming tasks are listed in the first list,
enhance task scheduling and resource allocation in a starting with the first one. The following two lists are
fog setting. The suggested system comprises a cloud used to filter this list. The 3rd list contains the tasks
layer, a fog layer, and an integrated fog environment which have been canceled by the users and is known
made up of each of these layers. The base broker is in as the urban list such as the tasks which have been not
charge of arranging and assigning jobs to the proper treated. The 2nd list consists of the tasks that would
fog nodes after receiving user requests from the fog be dispatched to the contract while it waited for the
nodes. The cloud layer handles intricate tasks. The broker to choose a free node.
suggested method attempts to shorten latency and The Levy flight is used by the method to randomly
speed up task execution. place tasks (eggs) on the fog nodes that make up
the search space. While some eggs may develop into
Discussion about the proposed method superior solutions, others might not. The program
To organize and schedule tasks in the environment employs the process of a local search to fine-tune
of fog computing, the study suggests a queuing para- the solutions obtained and the random search of the
digm. Three lists make up the model: a block list for cuckoos to locate new places to deposit their eggs.
the tasks, a schedule list for incoming tasks that have While the local search process aids in enhancing the
been canceled or are not essential, as well as a wait- quality of these solutions, the Levy flight in CS offers
ing list for the tasks which must wait for free nodes. an effective approach to scouring the search space for
To identify whether a node is full and to select the fresh answers. The following steps are used to put this
nearest node for processing of task, the model com- suggestion into practice:
putes loaded factor and distance. The task scheduling
method to appropriate nodes is then optimized using • Initialization: Create a random beginning popu-
the CS algorithm. For putting the algorithm into lation of potential solutions. Have a colony of
practice and adding it to a network of fog comput- cuckoos that utilize the Levy flight to randomly
ing, the article offers four guidelines. As illustrated in lay their eggs on the searched space.
Figure 35.1, the nodes process the tasks after which • Fitness evaluation: To assess each potential solu-
the users receive the results. The primary node (the tion’s suitability, and evaluate its objective func-
broker) is considered the system’s operational core in tion. Utilizing Equation (2), determine the loaded
this form. According to the idea, the broker organizes factor.
Applied Data Science and Smart Systems 261
(2)
(3)
Figure 35.2 A flowchart of the proposal’s steps
where, (x1, y1, z1) and (x2, y2, z2) indicate the coordi-
nates of the 2 points.
You must first subtract the respective coordinates
from these two positions, square each variation, add Evaluation measurement
the squared variations, and then take the square root In fog computing, latency is the time amount that
of the sum. passes between the process started and its accom-
This computes as: (x1 – x2)2, (y1 – y2)2 and (z1 – z2)2 plishment. For many applications, lowering latency
signifies the squared difference among the x, y and is a key objective since it may have a big influence
z-coordinates of 2 positions by computing Equation on how well fog computing systems function. The
(2 and 3) could examine the available and closet node latency in Equation (4) will be calculated.
as a better node.
Sort the tasks: employing a three-list proposal: (4)
Block, wait, and scheduling list, which represents
requests that have been refused and are waiting for P indicates the time of processing; N denotes the
an available node. delay of a network.
Selection of nest: Replace a nest (such as a candi- The P may be estimated by multiplying the R
date solution) with a new candidate solution if the old (processing rate) by the processing time (T) for the
candidate solution has a low fitness value. task:
Generating new solutions: Choose between a ran-
dom flight or random walk a to get a new candidate P=R×T (5)
solution. The generation is carried out under the 3
lists’ suggested mechanisms. The transmission time (Tt) and the propagation time
Acceptance criterion: If the novel candidate solu- (Tp) might be added to examine the network delay:
tion’s fitness value is greater as compared to the current
solution, it will be approved. If not, it is accepted by a
N=Tt + Tp(6)
specific probability depending on the variation in fit-
ness values between the recent and previous solutions. The transmission time could be computed as the
Stopping criterion: Run the method until a stop- product of transmission rate (Rt) along with the size
ping requirement is satisfied like when the several of the task (S):
iterations have attained a certain limit or the fitness
value is adequate.
Return the best solution: Return the optimal candidate (7)
solution discovered as an optimization consequence.
Here, created a flow chart for this suggested pro- When the propagation speed (S) and distance (d) are
cess, as seen in Figure 35.2. multiplied, the propagation time may be found.
262 Enhancing latency performance in fog computing through intelligent resource allocation
Table 35.2 Parameters of the fog environment a resource in fog computing, especially in mitigat-
ing latency-related difficulties. This success can be
Parameter Value
attributed to its unique features: combining distance
Router 6 measurements between fog nodes and users into the
fitness function and applying queue scheduling with
Users’ node 33
3 suggested lists in the basic broker, all of which are
Fog node 12 dependent on loaded criteria to decide whether or
Base broker 3 not a node is available. This precision in node selec-
tion significantly reduces randomness, enhancing the
overall system performance. Latency, a pivotal metric
Table 35.3 Final outcome
in assessing resource allocation, encompasses vari-
ous phases of request processing, from reception and
Latency (µs) Latency (µs) Distance Distance proposal scheduling to response transmission. The
with Cuckoo without (km) with (km) without suggested approach, which is founded on the CS algo-
Cuckoo Cuckoo Cuckoo rithm, exhibits significant benefits in terms of cutting
down on latency and increasing system efficacy. These
0.0092 0.0284 7.0711 21.21324
results highlight the utmost significance of the CS
0.0130 0.0748 5.02 28.0180 algorithm and its potentially game-changing role in
0.0064 0.0068 1.01 1.02 fog computing system optimization. Future research
0.0213 0.1086 7.0712 35.3555 endeavors could delve deeper, focusing on further
advancements and optimizations, especially in the
realm of IoT applications where the CS method holds
significant promise for resource allocation.
Evaluation and research in experiments
The CS method is proposed in the work as a way to References
reduce latency and delay across small distances. The
Abd-Ali, R. S., Radhi, S. A., and Rasool, Z. I. (2020). A sur-
strategy entails choosing the ideal set of the param-
vey: The role of the internet of things in the develop-
eters listed in Table 35.2. This could reduce latency ment of education. Indonesian J. Elect. Engg. Comp.
and boost system efficiency. Python is used for the Sci. 19(1), 215–221. [Link]
implementation of the visual studio code. v19.i1.pp215-221.
The suggested method reduces latency by determin- Abd Elaziz, M., Abualigah, L., and Attiya, I. (2021). Ad-
ing the closest node to each user using the CS algo- vanced optimization technique for scheduling IoT
rithm and estimating the distance between nodes and tasks in cloud-fog computing environments. Fut. Gen.
users. The program successfully identified the occa- Comp. Sys., 124, 142–154.
sion nodes for processing 4 user requests delivered to Agarwal, M. and Srivastava, G. M. S. (2018). A Cuckoo
a base broker during an evaluation. The findings illus- search algorithm-based task scheduling in cloud
computing. Adv. Intel. Sys. Comput., 554, 293–299.
trate a substantial drop in latency when employing the
[Link]
CS method. This method was utilized to evaluate its
Alsadie, D. (2022). Resource management strategies in fog
effectiveness by comparing latency with and without computing environment-A comprehensive review. Int.
the method. The table contrasts the delay in microsec- J. Comp. Sci. Netw. Sec., 22(4), 310–328.
onds and the distance in kilometers while using and Bhatia, H., Panda, S. N., and Nagpal, D. (2020). Internet of
without the cuckoo algorithm. As depicted in Table things and its applications in healthcare-A survey. 2020
35.3, the statistics prominently underscore the con- 8th Int. Conf. Reliab. Infocom Technol. Optim. (Trends
siderable reduction in distance achieved through the and Future Directions) (ICRITO), 305–310. https://
cuckoo algorithm implementation. This reduction in [Link]/10.1109/ICRITO48877.2020.9197816.
distance is assessed by computing latency and mea- Bittencourt, L. F., Diaz-Montes, J., Buyya, R., Rana, O. F.,
suring the distance among work scenarios with as and Parashar, M. (2017). Mobility-Aware application
scheduling in fog computing. IEEE Cloud Comput.,
well as without the cuckoo algorithm. The analysis
4(2), 26–35. [Link]
discerns the disparity between these results and effec-
Ghobaei-Arani, M., Souri, A., and Rahmanian, A. A. (2020).
tively illustrates the performance enhancement attrib- Resource management approaches in fog computing:
uted to the cuckoo algorithm. A comprehensive review. J. Grid Comput., 18(1),
1–42.
Conclusion Hao, Z., Ed Novak, Yi, S., and Li, Q. (2017). Challenges
and software architecture for fog computing. IEEE
In summary, the CS algorithm has been developed Int. Comput., 21(2), 44–53. [Link]
as a highly efficient technique for the allocation of MIC.2017.26.
Applied Data Science and Smart Systems 263
Hong, Cheol-ho, and Varghese, B. (2018). Resource manage- ronment. Indonesian J. Elec. Engg. Comp. Sci., 18(2),
ment in fog / edge computing: A survey. J. Supercom- 1081–1088. [Link]
put. 75. [Link] i2.pp1081-1088.
Islam, M. S. Ul, Kumar, A., and Hu, Y.-C. (2021). Context- Rani, M., Guleria, K., and Panda, S. N. (2021a). Cloud com-
aware scheduling in fog computing: A survey, taxono- puting: An empowering technology: architecture, ap-
my, challenges, and future directions. J. Netw. Comp. plications and challenges. 2021 9th Int. Conf. Reliab.
Appl., 180, 103008. Infocom Technol. Optim. (Trends and Future Direc-
Iyapparaja, M., Naif Khalaf Alshammari, M. Sathish Ku- tions), ICRITO 2021, 1–6. [Link]
mar, S. Krishnan, and Chiranji Lal Chowdhary. (2022). ICRITO51393.2021.9596259.
Efficient Resource Allocation in Fog Computing Us- Rani, M., Guleria, K., and Panda, S. N. (2021b). Enhancing
ing QTCS Model. Computers, Materials & Continua. performance of cloud: Fog computing architecture,
70(2). DOI:10.32604/cmc.2022.015707, 1–15. challenges and open issues. 2021 9th Int. Conf. Reliab.
Jamil, B., Ijaz, H., Shojafar, M., Munir, K., and Buyya, R. Infocom Technol. Optim. (Trends and Future Direc-
(2022). Resource allocation and task scheduling in fog tions)(ICRITO), 1–7.
computing and internet of everything environments: A Seth, Ishita, Kalpna Guleria, and Surya Narayan Panda.
taxonomy, review, and future directions. ACM Com- (2023). A lane-based advanced forwarding protocol
put. Sur., 54(11s). [Link] for internet of vehicles. International Journal of Per-
Jamil, Bushra, Humaira Ijaz, Mohammad Shojafar, Kashif vasive Computing and Communications. [Link]
Munir, and Rajkumar Buyya. (2020). Resource alloca- org/10.1108/IJPCC-08-2022-0305
tion and task scheduling in fog computing and internet Shamman, A. H., Alasadi, H. A., Ameen, H. A., and Rasol,
of everything environments: A taxonomy, review, and Z. I. (2022). Cost-effective resource and task schedul-
future directions. ACM Computing Surveys (CSUR). ing in fog nodes cost-effective resource and task sched-
54(11s): 1–38. uling in fog nodes. 466–477. [Link]
Kaur, Mandeep, and Rajni Aron. (2021). A systematic study ijeecs.v27.i1.pp466-477.
of load balancing approaches in the fog computing Singh, P. and Singh, R. (2022). Energy-efficient delay-aware
environment. The Journal of supercomputing. 77(8), task offloading in fog-cloud computing system for IoT
9202–9247. sensor applications. J. Netw. Sys. Manag., 30(1), 1–25.
Khattar, N., Sidhu, J., and Singh, J. (2019). Toward energy- [Link]
efficient cloud computing: A survey of dynamic power Tran-Dang, H., Bhardwaj, S., Rahim, T., Musaddiq, A.,
management and heuristics-based optimization tech- and Kim, D.-S. (2022). Reinforcement learning based
niques. J. Supercomput., 75. [Link] resource management for fog computing environ-
s11227-019-02764-2. ment: Literature review, challenges, and open issues. J.
Ma, K., Bagula, A., Nyirenda, C., and Ajayi, O. (2019). An Comm. Netw., 24(1), 83–98. [Link]
Iot-based fog computing model. Sensors (Switzerland), jcn.2021.000041.
19(12), 1–17. [Link] Vashisht, P. and Kumar, V. (2022). A cost effective and en-
Mahmud, Redowan, Kotagiri Ramamohanarao, and Rajku- ergy efficient algorithm for cloud computing. Int. J.
mar Buyya. (2020). Application management in fog com- Math. Engg. Manag. Sci., 7(5), 681–696. [Link]
puting environments: A taxonomy, review and future di- org/10.33889/IJMEMS.2022.7.5.045.
rections. ACM Computing Surveys (CSUR). 53(4). 1–43. Wadhwa, H. and Aron, R. (2022). TRAM: Technique for re-
Martinez, I., Hafid, A. S., and Jarray, A. (2021). Design, re- source allocation and management in fog computing
source management, and evaluation of fog computing environment. J. Supercomput., 78(1), 667–690.
systems: A survey. IEEE Int. Things J., 8(4), 2494– Yin, C., Li, T., Qu, X., and Yuan, S. (2020). An improved
2516. [Link] ant colony optimization job scheduling algorithm in
Member, Student, and Luo, J. (2018). Tasks scheduling fog computing. Int. Symp. Artif. Intel. Robotics 2020,
and resource allocation in fog computing based on 11574, 132–141.
containers for smart manufacturing. IEEE Trans. Yousefpour, A., Fung, C., Nguyen, T., Kadiyala, K., Jalali,
Indus. Informat., 14(10), 4712–4721. [Link] F., Niakanlahiji, A., Kong, J., and Jue, J. P. (2019).
org/10.1109/TII.2018.2851241. All one needs to know about fog computing and re-
Moshref, Mahmoud, Rizik Al-Sayyed, and Saleh Al- lated edge computing paradigms: A complete survey.
Sharaeh. (2022). Improving the quality of service in J. Sys. Arch., 98, 289–330. [Link]
wireless sensor networks using an enhanced routing sysarc.2019.02.009.
genetic protocol for four objectives. Indonesian Jour- Zhang, L. and Li, J. (2018). Enabling robust and privacy-
nal of Electrical Engineering and Computer Science. preserving resource allocation in fog computing. IEEE
26(2): 1182–1196. Acc., 6, 50384–50393. [Link]
Nazir, S., Shafiq, S., Iqbal, Z., Zeeshan, M., Tariq, S., and CESS.2018.2868920.
Javaid, N. (2019). Cuckoo optimization algorithm Zhou, Zhenyu, Pengju Liu, Junhao Feng, Yan Zhang, Sha-
based job scheduling using cloud and fog computing hid Mumtaz, and Jonathan Rodriguez. (2019). Com-
in smart grid. Adv. Intel. Netw. Collab. Sys. 10th Int. putation resource allocation and task assignment
Conf. Intel. Netw. Collab. Sys. (INCoS-2018), 34–46. optimization in vehicular fog computing: A contract-
Potluri, S. and Rao, K. S. (2020). Optimization model for matching approach. IEEE Transactions on Vehicular
QoS based task scheduling in cloud computing envi- Technology. 68(4): 3113–3125.
36 Pediatric thyroid ultrasound image classification using
deep learning: A review
Jatinder Kumar1,a, Surya Narayan Panda1 and Devi Dayal2
1
Department of Computer Science and Engineering, Chitkara University, Punjab, India
2
Department of Paediatrics, Endocrinology and Diabetes Unit, PGIMER, Chandigarh, India
Abstract
The thyroid gland, a little butterfly-shaped gland at the front of the neck, generates hormones which govern metabolism.
Thyroid problems are most typically detected and classified via ultrasound (US) imaging. US imaging has become one of
the most important contributions for analyzing thyroid disorders due to its safety, accessibility, non-invasiveness and cost-
effectiveness. Machine learning (ML) advances, especially deep learning (DL) is proving to be beneficial in recognizing and
quantifying patterns in clinical images. At the heart of these advancements is DL algorithms’ ability to extract hierarchical
feature representations directly from images, eliminating the requirement for constructed features. This study describes the
evolution of ML, the concepts of DL algorithms, and an overview of successful applications, including clinical picture seg-
mentation for US imaging of thyroid-related illnesses. Finally, certain research difficulties are mentioned along with future
enhancements.
[Link]@[Link]
Applied Data Science and Smart Systems 265
Table 36.1 Thyroid diseases and their symptoms
Semi-automatic segmentation strategies utilizing performance times, along with access to huge datasets
automated algorithms require a minimal amount and progressions in learning methods, have supported
of user input to get effective segmentation results the rise in the use of deep learning systems (Shen et
(Iglesias et al., 2017; Gera et al., 2021). The user might al., 2017).
be prompted to select an approximate initial ROI The forthcoming outline provides an overview of
that will serve as the basis for segmenting the entire the review’s structure. The following section delves
image. It may be necessary to do physical verification into AI and its related technological methodolo-
and remove region margins in order to diminish seg- gies, exploring ML role in image segmentation. DL
mentation fault. Techniques for semiautomatic seg- techniques for clinical photo segmentation and their
mentation comprise (1) seeded region growth (SRG) architectures, common methods of implementing DL
method, which combines neighboring pixels with like architectures, and metrics for evaluating image seg-
intensities iteratively established on a user-supplied mentation performance. Recent applications of DL
first seed idea; (2) Iteratively altering initial boundary models in various biological image segmentation con-
forms represented by contours utilizing a shrinkage or texts are examined following the above section. Deep
expanding procedure established on the implied level learning architecture implementation methodologies
of a utility using a level set established active contour were discussed. Literature review was done and in
model, which has the advantage of requiring no prior the end discussions concerning the challenges related
shape information or initial ROI locations and (3) with DL based image segmentation, the concluding
restricted area based dynamic contour approaches, remarks, and potential avenues for further research.
which use area parameters to characterize the image’s
foreground and background using small local regions Deep learning overview
and can handle heterogeneous textures (Zhang et al.,
2012; Fan et al., 2015; Kim et al., 2016). Artificial intelligence (AI)
The user is not required to interact with the com- In general, AI is described as the use of any equip-
pletely automated segmentation procedures. Shape ment to simulate the human cognitive process, which
models, atlas-defined division techniques, random includes learning, applying, and solving difficult prob-
forests, and deep neural networks are all supervised lems. Figure 36.7 shows the hierarchical links between
learning procedures that need drill information. AI, an area of computer science that comprises ML,
Unsupervised learning techniques require labeled DL, and convolutional neural networks (CNNs). AI,
pictures generated through manual segmentation dubbed “the fourth industrial revolution”, is signifi-
for both training and validation data, incurring the cantly transforming the terrain of our entire lives
same constraints as previously mentioned. Limited
contrast between regions and the significant varia-
tions in forms, sizes, textures, and colors of the ROI
further introduce challenges in the computerized seg-
mentation of clinical pictures (Roth et al., 2018). Big
disparities in the resource photos data might result
from noise in the acquisition of source data, which is
prevalent in real-world applications. As a result, max-
imum current systems established through clustering
methods, watershed procedures, and machine learn-
ing established methodologies for a fundamental lack
in the worldwide application, limiting their usage to a
minor quantity of applications. Furthermore, the prac-
tice of social feature exchange, often employed along-
side ML methods based on support vector machines
(SVM) or neural networks (NN), is ineffective, inca-
pable of handling novel data in its recent application,
and does not typically adjust to newly introduced
information. Deep learning algorithms may be able
to process raw data without the requirement for pre-
defined features. Natural image segmentation for
semantic reasons, as well as biological image segmen-
tation, has all been effectively accomplished using
these methods (Fujita et al., 2018). Quicker CPUs and Figure 36.7 Hierarchical relationships of AI, ML, DL
GPUs, which substantially concentrated exercise, and and CNN
Applied Data Science and Smart Systems 269
today. The phrase “artificial intelligence” was initially In order to categorize ROI as healthy or diseased, a
proposed during a symposium at Dartmouth in 1956 common approach involves employing ML for image
and its evolution is shown in Figure 36.8. Importantly, segmentation. The initial phase of constructing such
AI approaches are highly suited to imaging-based an application involves pre-processing, which may
domains since the image itself is the primary source encompass noise removal or contrast enhancement
of data for training AI algorithms because pixel val- through filters. After pre-processing, the image under-
ues can be quantified (Russell et al., 2010; Yoon et al., goes segmentation via methods like thresholding,
2017; Dhiman et al., 2022). clustering, and edge-based segmentation. Once seg-
mented, attributes related to color, texture, contrast,
Machine learning (ML) and size are extracted from the ROI. Subsequently,
ML established the picture dissection method which using feature selection techniques like principal com-
is frequently used for categorizing ROI, such as ponent analysis (PCA) or statistical analysis, signifi-
unhealthy or healthy regions. Pre-processing, which cant attributes are identified. These chosen features
can include the usage of a filter to eliminate blare are then fed into a ML classifier, such as SVM or NN.
or for distinction improvement, is the initial step in The ML classifier determines optimal boundaries
constructing such an application. The image is seg- between classes by integrating the input feature vec-
mented after it has been pre-processed, utilizing tor with respective class labels (Kumar et al., 2021).
techniques such as thresholding, clustering, and edge- The ML classifier can then be used to categorize
based segmentation. Color, texture, dissimilarity, and fresh data. Common factors include addressing neces-
dimension features are extracted from the ROI after sary pre-processing requirements for raw picture data,
segmentation. figuring out the right features and the dimensions of
270 Pediatric thyroid ultrasound image classification using deep learning: A review
minimizing radiation exposure by using appropriate biological tissues using light waves. It is most typically
techniques and protocols. used in ophthalmology to see and analyze the retina
and other ocular components, but it is also utilized in
Ultrasound imaging (US) cardiology and dermatology. OCT functions based on
US is a commonly employed method for diagnosing the principle of interferometry, wherein a light beam
and providing real-time visualization of organs, tis- is divided into two distinct paths: a reference path and
sues, and blood flow. In this technique, a handheld a sample path. The reference beam is directed towards
transducer device is positioned on the skin’s surface, a mirror, while the sample beam is directed towards
emitting sound waves into the body. These high- the tissue being analyzed. Subsequently, light waves
frequency sound waves, beyond the range of human originating from both directions engage with the tis-
hearing, traverse through the body and rebound sue before retracing their path to the apparatus.
upon encountering diverse tissues and structures, In ophthalmology, OCT is particularly valuable for
offering valuable insights during ultrasound exami- visualizing the retina and identifying various retinal
nations. Transducer also acts as a receiver, capturing conditions, such as macular degeneration, diabetic ret-
the reflected sound waves. The reflected sound waves inopathy, and glaucoma. It allows clinicians to assess
are then processed by a computer to create real-time the thickness, integrity, and pathological changes in
images on a monitor. These images depict the shape, retinal layers, aiding in diagnosis, treatment planning,
size, and composition of organs, tissues, and structures and monitoring of these conditions.
within the body. The ultrasound images are typically OCT is also used in other medical specialties. In
gray-scale, but color Doppler can be used to visualize cardiology, it can provide detailed imaging of blood
and assess blood [Link] imaging is frequently uti- vessels and identify atherosclerotic plaques in coro-
lized in obstetrics and gynecology, cardiology, radi- nary arteries. In dermatology, OCT can help visualize
ology, and abdominal imaging, among other medical and assess skin lesions and guide biopsies. One of the
disciplines. It can provide important details regarding key advantages of OCT is its high-resolution imaging
the structure and function of organs such the heart, capability, allowing for the visualization of fine tissue
liver, kidneys, and reproductive organs. In obstetrics, structures with micrometer-level precision. However,
it is commonly used for monitoring fetal develop- OCT has certain limitations. It is primarily limited to
ment during pregnancy. One of the key advantages of imaging superficial tissues due to the limited penetra-
ultrasound imaging is its safety and non-invasiveness. tion depth of light waves. Medical images taken at
While ultrasonic imaging offers numerous benefits, it a microscopic level are utilized to assess the tissue’s
also has some drawbacks (Reddy et al., 2008; Gharib small structure. The biopsy is utilized to retrieve tis-
et al., 2010; Haugen et al., 2016; Haque et al., 2020). sue for investigation, and subsequently, staining com-
ponents are employed to expose cellular features in
Magnetic resonance imaging (MRI) areas of the tissue. Counter stains are used to give the
The transmitted signals are picked up by a collec- graphics more color, visibility, and contrast.
tion of specialized antennas in the MRI machine,
and this information is processed by a computer to Data augmentation
create comprehensive cross-sectional images of the
body. These images, which can be viewed from dif- The enactment of DL neural networks was determined
ferent angles, provide information about the structure by the handiness of appropriate facts. Facts augmen-
and composition of various tissues and organs. It can tation, which includes removing a set of reasonable
help identify abnormalities, such as tumors, inflam- alterations to the samples (e.g., flip, rotate, mirror) in
mation, or structural abnormalities, and assist in the addition to augmenting color (grey) values, is the great-
diagnosis and monitoring of conditions like strokes, est commonly use method for aggregate the extent of
multiple sclerosis, joint disorders, and certain types of the training dataset. The efficiency of data augmen-
cancer. One notable advantage of MRI is that it does tation is examined in non-clinical research, and the
not require ionizing radiation exposure, making it a outcomes tell that classic augmentation methods can
safe imaging alternative. However, certain contrain- enrich by up to 7% (Fotenos et al., 2005; Golan et al.,
dications, such as the presence of metallic implants 2016; Milletari et al., 2017). When real data is scarce,
or devices in the body, may restrict its use in some numerous data augmentation procedures are used to
individuals (Kwak et al., 2011). generate more training data from the existing dataset.
These augmentation techniques change images while
Optical coherence tomography (OCT) and micro- retaining their class, and they can include methods
scopic images and scintigraphy such as – (1) Image translation: This technique involves
OCT is a non-invasive medical imaging technology shifting the pixels of an image along a single direc-
that captures high-resolution cross-sectional images of tion, either horizontally or vertically, without altering
274 Pediatric thyroid ultrasound image classification using deep learning: A review
is essential as the enduring advantages of data sharing variability in manual spot segmentation and its effect
far surpass any transient gains attained through data on spot quantitation in two-dimensional electropho-
concealment. DL algorithms have ushered in unpar- resis analysis. Electrophoresis, 31(10), 1739–1742.
alleled enhancements in performance across diverse Iglesias, J. E. (2017). Globally optimal coupled surfaces for
semi-automatic segmentation of medical images. Int.
healthcare domains, spanning from the automated
Conf. Inform. Proc. Med. Imag., 610–621.
segmentation of CT images to the analysis of thyroid
Fan, M. and Lee, T. C. M. (2015). Variants of seeded region
US images. However, there is further potential to be growing. IET Image Proc. 9(6), 478–485.
realized through the augmentation of publicly acces- Zhang, H., Albert, M., and Willig, A. (2012). Combining
sible labeled images. The manual annotation of visual TDMA with slotted Aloha for delay constrained traf-
data by experts remains a notable impediment in gen- fic over lossy links. 2012 12th Int. Conf. Con. Au-
erating accurate ground truths. In instances where tomat. Robot. Vis. (ICARCV), 701–706.
ground truth is absent, greater emphasis should be Kim, Y. J., Lee, S. H., Park, C. M., and Kim, K. G. (2016).
placed on unsupervised learning strategies. Evaluation of semi-automatic segmentation methods
for persistent ground glass nodules on thin-section CT
scans. Healthcare Informat. Res., 22(4), 305–315.
References Roth, H. R., Shen, C., Oda, H., Oda, M., Hayashi, Y., Misa-
Er, O., Sertkaya, C., Temurtas, F., and Tanrikulu, A. C. wa, K., and Mori, K. (2018). Deep learning and its ap-
(2009). A comparative study on chronic obstructive plication to medical image segmentation. Med. Imag.
pulmonary and pneumonia diseases diagnosis using Technol., 36(2), 63–71.
neural networks and artificial immune system. J. Med. Fujita, Hiroshi, Takeshi Hara, Xiangrong Zhou, Kagaku
Sys., 33, 485–492. Azuma, Daisuke Fukuoka, Yuji Hatanaka, Naoki
Vaz, V. A. S. (2014). Diagnosis of hypo and hyperthyroid Kamiya et al. A02-3 Function Integrated Diagnostic
using MLPN network. Int. J. Innov. Res. Sci. Engg. Assistance Based on Multidisciplinary Computational
Technol., 3(7), 14314–14323. Anatomy Models. The 5th International Symposium
Rivkees, S. and Bauer, A. J. (2021). Thyroid disorders in on Multidisciplinary Computational Anatomy, 1–13.
children and adolescents. Ped. Endocrinol., 395–424. Shen, D., Wu, G., and Heung-Il Suk. (2017). Deep learning
Dayal, D., Prasad, R., Bhunwal, S., Kumar, R., Kumar, R. in medical image analysis. Ann. Rev. Biomed. Engg.,
M., and Sodhi, K. S. (2017). Spectrum of extrathyroi- 19, 221–248.
dal congenital malformations in a cohort of North Yoon, D. (2017). What we need to prepare for the fourth in-
Indian children with permanent primary congenital dustrial revolution. Healthcare Informat. Res., 23(2),
hypothyroidism. Thyroid Res. Prac. 14(1), 8–11. 75–76.
Dayal, D. and Gupta, B. M. (2020). Pediatric hyperthy- Russell, Stuart J., and Peter Norvig. (2010). Artificial intel-
roidism research: A scientometric assessment of ligence a modern approach. London, 2010.
global publications during 1990–2019. Thyroid Res. Dhiman, P., Kukreja, V., Manoharan, P., Kaur, A., Kamruz-
Prac.,17(3), 134–140. zaman, M. M., Dhaou, I. B., and Iwendi, C. (2022). A
Vasavi, J., and M. S. Abirami. (2020). A qualitative perfor- novel deep learning model for detection of severity lev-
mance comparison of supervised machine learning el of the disease in citrus fruits. Electronics, 11(3), 495.
algorithms for iris recognition. European Journal of Kumar, A., Sharma, S., Goyal, N., Singh, A., Cheng, X., and
Molecular & Clinical Medicine. 7(6): 2020. Singh, P. (2021). Secure and energy-efficient smart
Masood, S., Sharif, M., Raza, M., Yasmin, M., Iqbal, M., building architecture with emerging technology IoT.
and Javed, M. Y. (2015). Glaucoma disease: A survey. Comp. Comm., 176, 207–217.
Cur. Med. Imag., 11(4), 272–283. Gera, T., Singh, J., Mehbodniya, A., Webber, J. L., Shabaz,
Pal, A., Chaturvedi, A., Garain, U., Chandra, A., and Chat- M., Thakur, D. (2021). Dominant feature selection
terjee, R. (2016). Severity grading of psoriatic plaques and machine learning-based hybrid approach to ana-
using deep CNN based multi-task learning. 2016 23rd lyze android ransomware. Sec. Comm. Netw., 1–22.
Int. Conf. Pat. Recogn. (ICPR), 1478–1483 Deng, L. and Yu, D. (2014). Deep learning: Methods and
Wang, Ge. (2016). A perspective on deep imaging. IEEE applications. Foundat. Trends Sig. Proc., 7(3–4), 197–
Acc., 4, 8914–8924. 387.
Silva, Flávio Henrique Schuindt da. (2018). Deep learning Suzuki, K. (2017). Overview of deep learning in medical
for corpus callosum segmentation in brain magnetic imaging. Radiol. Phys. Technol., 10(3), 257–273.
resonance images. Universidade Federal do Rio de Ja- Krizhevsky, A. (2012). Advances in neural information pro-
neiro, 1–122. cessing systems. (No Title), 1097.
Volkenandt, T., Freitag, S., and Rauscher, M. (2018). Ma- Garcia-Garcia, A., Orts-Escolano, S., Oprea, S., Villena-
chine learning powered image segmentation. Mi- Martinez, V., Martinez-Gonzalez, P., and Garcia-
croscop. Microanal., 24(S1), 520–521. Rodriguez, J. (2018). A survey on deep learning tech-
Iþýn, A., Direkoðlu, C., and ªah, M. (2016). Review of MRI- niques for image and video semantic segmentation.
based brain tumor image segmentation using deep Appl. Soft Comput., 70, 41–65.
learning methods. Proc. Comp. Sci., 102, 317–324. Ronneberger, O., Fischer, P., and Brox, T. (2015). U-net:
Millioni, R., Sbrignadello, S., Tura, A., Iori, E., Murphy, E., Convolutional networks for biomedical image segmen-
and Tessari, P. (2010). The interand intra-operator tation. Med. Image Comput. Computer-Assisted In-
Applied Data Science and Smart Systems 277
terven.–MICCAI 2015: 18th Int. Conf. Munich, Ger- Perez, L. and Wang, J. (2017). The effectiveness of data aug-
many, October 5-9, 2015, Proc. Part III 18, 234–241. mentation in image classification using deep learning.
Milletari, F., Navab, N., and Ahmadi, S.-A. (2016) V-net: arXiv preprint arXiv:1712.04621.
Fully convolutional neural networks for volumetric Shie, C.-K., Chuang, C.-H., Chou, C.-H., Wu, M.-H., and
medical image segmentation. 2016 Fourth Int. Conf. Chang, E. Y. (2015). Transfer representation learning
3D Vis. (3DV), 565–571. for medical image analysis. 2015 37th Ann. Int. Conf.
Davamani, K. A., Rene Robin, C. R., Amudha, S., and Jani IEEE Engg. Med. Biol. Soc. (EMBC), 711–714.
Anbarasi, L. (2021). Biomedical image segmentation Lalkhen, A. G. and McCluskey, A. (2008). Clinical tests:
by deep learning methods. Computat. Anal. Deep sensitivity and specificity. Cont. Educ. Anaes. Crit.
Learn. Med. Care Prin. Meth. Appl., 131–154. Care Pain, 8(6), 221–223.
Haque, I. R. I. and Neubert, J. (2020). Deep learning ap- Van S., Karlijn J., Stel, V. S., Reitsma, J. B., Dekker, F. W.,
proaches to biomedical image segmentation. Infor- Zoccali, C., and Jager, K. J. (2009). Diagnostic meth-
mat. Med. Unlocked, 18, 100297. ods I: Sensitivity, specificity, and other measures of ac-
Reddy, U. M., Filly, R. A., and Copel, J. A. (2008). Prena- curacy. Kidney Int., 75(12), 1257–1263.
tal imaging: ultrasonography and magnetic resonance Csurka, G., Larlus, D., Perronnin, F., and Meylan, F. (2013).
imaging. Obstet. Gynecol., 112(1), 145. What is a good evaluation measure for semantic seg-
Haugen, B. R., Alexander, E. K., Bible, K. C., Doherty, G. mentation? BMVC, 27, 10–5244.
M., Mandel, S. J., Nikiforov, Y. E., Pacini, F. et al. Wong, H. B. and Lim, G. H. (2011). Measures of diagnostic
(2016). 2015 American Thyroid Association manage- accuracy: Sensitivity, specificity, PPV and NPV. Proc.
ment guidelines for adult patients with thyroid nod- Singapore Healthcare, 20(4), 316–318.
ules and differentiated thyroid cancer: the American Xu, Y., Wang, Y., Yuan, J., Cheng, Q., Wang, X., and Carson,
Thyroid Association guidelines task force on thyroid P. L. (2019). Medical breast ultrasound image segmen-
nodules and differentiated thyroid cancer. Thyroid, tation by machine learning. Ultrasonics, 91, 1–9.
26(1), 1–133. Badea, M.-S., Felea, I.-I., Florea, L. M., and Vertan, C.
Thakur, D., Singh, J., Dhiman, G., Shabaz, M., and Gera, (2016). The use of deep learning in image segmenta-
T. (2021). Identifying major research areas and minor tion, classification and detection. arXiv preprint arX-
research themes of android malware analysis and de- iv:1605.09612.
tection field using LSA. Complexity, 1–28. Kaur, Jaspreet, and Alka Jindal. (2012). Comparison of thy-
Gharib, H., Papini, E., Paschke, R., Duick, D. S., Valcavi, roid segmentation algorithms in ultrasound and scin-
R., Hegedüs, L., Vitti, P., and AACE/AME/ETA Task tigraphy images. International Journal of Computer
Force on Thyroid Nodules. (2010). American Associa- Applications. 50(23), 1–4.
tion of Clinical Endocrinologists, Associazione Medici Poudel, Prabal, Alfredo Illanes, Debdoot Sheet, and
Endocrinologi, and European Thyroid Association Michael Friebe. (2018). Evaluation of commonly
medical guidelines for clinical practice for the diag- used algorithms for thyroid ultrasound images seg-
nosis and management of thyroid nodules: executive mentation and improvement using machine learn-
summary of recommendations. J. Endocrinol. Investi- ing approaches. Journal of healthcare engineering.
gat., 33, 287–291. 2018. doi: [Link]
Kwak, J. Y., Han, K. H., Yoon, J. H., Moon, H. J., Son, E. 1–14.
J., Park, S. H., Jung, H. K., Choi, J. S., Kim, B. M., Shenoy, N. R. and Jatti, A. (2021). Ultrasound image seg-
and Kim, E.-K. (2011). Thyroid imaging reporting and mentation through deep learning based improvised
data system for US features of nodules: a step in estab- U-Net. Indonesian J. Elec. Engg. Comp. Sci., 21(3),
lishing better stratification of cancer risk. Radiology, 1424–1434.
260(3), 892–899. Shah, Chintan, and Anjali G. Jivani. (2013). Comparison of
Park, J.-Y., Lee, H. J., Jang, H. W., Kim, H. K., Yi, J. H., data mining classification algorithms for breast can-
Lee, W., and Kim, S. H. (2009). A proposal for a thy- cer prediction. In 2013 Fourth international confer-
roid imaging reporting and data system for ultrasound ence on computing, communications and networking
features of thyroid carcinoma. Thyroid, 19(11), 1257– technologies (ICCCNT). 1–4. IEEE, 2013. 10.1109/
1264. ICCCNT.2013.6726477
Fotenos, A. F., Snyder, A. Z., Girton, L. E., Morris, J. C., and Frannita, E. L., Nugroho, H. A., Nugroho, A., and Ardi-
Buckner, R. L. (2005). Normative estimates of cross- yanto, I. (2018). Thyroid nodule classification based
sectional and longitudinal brain volume decline in ag- on characteristic of margin using geometric and sta-
ing and AD. Neurology, 64(6), 1032–1039. tistical features. 2018 2nd Int. Conf. Biomed. Engg.
Golan, R., Jacob, C., and Denzinger, J. (2016). Lung nod- (IBIOMED), 54–59.
ule detection in CT images using deep convolutional Ying, X., Yu, Z., Yu, R., Li, X., Yu, M., Zhao, M., and Liu,
neural networks. 2016 Int. Joint Conf. Neu. Netw. K. (2018). Thyroid nodule segmentation in ultrasound
(IJCNN), 243–250. images based on cascaded convolutional neural net-
Milletari, F., Ahmadi, S.-A., Kroll, C., Plate, A., Rozanski, work. Neural Inform. Proc. 25th Int. Conf. ICONIP
V., Maiostre, J., Levin, J. et al. (2017). Hough-CNN: 2018, Siem Reap, Cambodia, December 13–16, 2018,
Deep learning for segmentation of deep brain regions Proc., Part VI 25, 373–384.
in MRI and ultrasound. Comp. Vis. Image Under-
standing, 164, 92–102.
37 Hybrid security of EMI using edge-based steganography
and three-layered cryptography
Divya Sharmaa and Chander Prabha
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
To enhance the security and ensure privacy of a larger data set of electronic medical images (EMI) each of which varies in
properties while they are in storage, or before being transmitted, and accessed through real-time applications has become a
challenging issue. Stored EMI should be easily accessible anytime while ensuring secrecy and privacy. The proposed hybrid
method (PHM) is a combination of steganography with cryptography which ensures security while reducing the computa-
tional time so that they can secure EMI in real time. Initially, in PHM the EMI is hidden using edge-based steganography and
then applied with three-layered cryptography. The proposed hybrid method is implemented using MATLAB. The efficiency
metrics applied are: total time which combines steganography with encryption time, decrypt and de-steganography time
thus overall processing time, Peak Signal to Noise Ratio (PSNR), Mean Square Error (MSE), Kullback-Leibler Divergence
(KLD), Root Mean Square Error (RMSE), Bit Error Rate (BER), etc. This article aims to secure a larger data set of 5856
EMI images of varying dimensions sized 1.16 GB by implementing the PHM which is a combination of cryptography and
steganography. Further performance analysis demonstrates its efficiency and effectiveness in terms of reduced total process-
ing time, encryption, and decryption time. Therefore, PHM can be used by hospitals to enhance security and privacy while
providing real-time access to EMI. The PHM achieved is 0.99, R is 0.99, while better value for Kullback-Leibler Divergence
(KLD), Root Mean Square Error (RMSE), Bit Error Rate (BER), Universal Average Changed Intensity (UACI), Number of
Changing Pixel Rate (NPCR), etc. Hence proving its statistical relevance.
divya009sharma@[Link]
a
Applied Data Science and Smart Systems 279
Figure 37.3 Commonly used smart medical devices that generate EHR
280 Hybrid security of EMI using edge-based steganography and three-layered cryptography
cryptography method for EMI security and protec- size in bytes, dimensions, and belonging to different
tion from modification, theft or loss attacks while patient’s security is an important aspect as it needs
they reside on storage devices. to be enhanced further while maintaining the original
EMI properties (Shukla et al., 2021). The easy to use,
Problem statement hassle free, anytime access to EMI is provided, faster
The previous researchers have introduced different access, and device scalable access to patients, medi-
data security schemas to enhance the security of EMI. cal practitioners, etc. EMI are used in case of medical
However, the previous studies have not effectively emergencies thus they need to be reliable, accurate,
enhanced data security. Most of them fail to mention and accessible in real time. Therefore, EMI security
encryption and decryption time (Adnan and Ariffin, and privacy need to be enhanced without affecting its
2019; Singh et al., 2020; Ali et al., 2022; Prabha et features.
al., 2022; Parmar and Shah, 2023) which helps prove
the efficiency of the proposed algorithm. However, Research contribution
computational time needs to be addressed while stor- The major contributions that led to this research have
ing EMI as they are used for real-time applications. been listed below:
The computational time is the time involved in access-
ing our EMI which involves de-steganography and 1) Developing a hybrid of steganography with
decryption process for the proposed hybrid method. cryptography method which is lightweight and
The previously used research works implemented con- capable of processing a diverse and larger data
ventional techniques such as Advanced Encryption set of 5856 EMI images.
Standard which is susceptible to brute force attacks 2) To propose a hybrid method that enhances secu-
(Adnan and Ariffin, 2019), RSA, least significant bit rity and privacy of the EMI images while main-
(LSB) steganography (Adee and Mouratidis, 2022), taining its picture quality for future diagnosis
Two-Fish algorithm (Maata, Cordova, and Halibas, and analysis also reducing the access time for the
2020). Thus, leading to proposed hybrid method same.
(PHM) which combines steganography with cryptog- 3) Efficiency evaluation of the proposed hybrid
raphy on X-ray images. The X-ray images are hidden method by analyzing the PHM time for encryp-
one at a time into a normalized cover image Lena. tion with steganography, decryption with extrac-
Then three-layers of cryptography are applied to tion time have been tabulated and compared
stego-image which will further enhance the security with previous research work along with the sta-
of EMI. This hybrid method is a light-weight combi- tistical test values such as PSNR, MSE, RMSE,
nation and is found suitable for real-time applications R, etc.
where EMI can be stored, or before transmitting them
over a network, and accessed in real-time (Sharma et This article is further sub-divided into the following
al., 2021) through real time applications. sections – the next is literature review which tabulates
the current state of research work, followed by PHM
Research motivation which gives an understanding of the proposed hybrid
The current study focuses on reducing the size of the method, then results and discussion of the proposed
EMI which renders them useless for future referenc- technique based on computational time, Finally, the
ing by medical practitioners, researcher, and insurance section discuss the conclusion of the proposed PHM
agencies, etc. Recently, there has been an increase in method.
the number of EMI users. The security and privacy of
EMI have become increasingly challenging as no EMI
is the same, having varying properties, thus, they vary Literature review
in size and have unique features. Due to varying sizes The current articles that are studied during this
(such as dimensions, storage space, etc.) and unique research work are tabulated in the form of Table 37.1.
features (therefore each image belongs to a unique This tabulation is based on the research goal that led
person at a unique point in time) for the EMI images. to their research work, the results achieved, the tech-
Previous studies have focused on normalizing the EMI nique proposed, and the future scope of research.
to dimensions 256 × 256 (El-Shafai et al., 2022) and
512 × 512 (Akkasaligar and Biradar, 2020; Brar et al.,
2022) which renders them useless for future medical Proposed hybrid method (PHM)
referencing. As medical data needs to be detailed, clear, The security of EMI from network hackers is an
and accurate for correct diagnosis. Thus, such meth- important challenge that needs to be addressed. This
ods have harmed the key feature of the EMI. Thus, research work focuses on enhancing the security and
EMI of varying properties such as number of pixels, privacy of the EMI (Ali et al., 2022) while in storage.
Applied Data Science and Smart Systems 281
Table 37.1 Literature review of studied research articles
(Al Hamid et EMR which exists as big Medical data is securely accessed Elliptic curve cryptography with 3
al., 2017) data to be secured from data and stored by decoy technique party one-round authenticates key
theft attacks, and security which allows only authorized exchange
breaches, in the cloud using fog users access
computing
(Ali et al., To implement a deep learning Improves security, anomalies, Novel method on blockchain
2022) algorithm that securely and monitors a user’s behavior, allowing remote encryption for
searches the distributed better efficiency compared to users and upload of a distributed
blockchain-based database peer block chain models. This ledger. The proposed method can
using homographic encryption technique supports immutability, be enhanced by applying methods
for secure access and searching tamper resistance, and delivery of such as the classification method
of records (implemented secured data resulting in reduced
using smart contracts and security breaches
Hyperledger tools)
(Lin et al., To develop a technique that Satisfactory decryption Proposed a multi-layered
2023) protects the confidentiality, performance, promising convolution processing network
reliability, and increases capabilities to protect the data (MCPN) cryptography combined
availability of digital images confidentiality, data recovery, and with artificial intelligence (AI)
while be processed by online data availability of digital images for cancer disease detection while
applications increasing its applicability to IoT
and IoMT by combining with
discreet Fourier transform at the
physical layer of data transmission
between heterogenous devices
(Parmar and Integrating IoT nodes with Performance and cost-effective IoT blockchain light-weight
Shah, 2023) blockchain solution with less performance cryptographic (IBLWC) approach
overhead
(Mothi and Retain the quality of the iris Achieved an increase in the A hybrid of wavelet packet
Karthikeyan, image after data hiding quality of the image. High- transform (WPT) and advance
2019) security hybrid for more reliable encryption standard (AES)
and secure cryptography cryptography
(Georgieva- Protect cardiac database Showed effectiveness, security, Daubechies wavelet transform
Tsaneva, against unauthorized access stability, and potential use in then conducted energy packing
Bogdanova, telemedicine efficiency-based compression
and
Gospodinova,
2022)
(Kumar et al., Transferring images over the Gives greater security, enhanced LSB steganography and AES
2022) Internet would face various and robust security, challenging cryptography
issues such as protection, to break by unauthorized access
copyrights, modification,
authentication
(Krishna, Security of data stored on Lesser time for implementing Fully homomorphic encryption
2018) a cloud. Asymmetric block cubic spline curve cryptography of Big Data using cubic spline
cipher mode used with global compared with error correction curve public key cryptography.
variable used for calculating code (ECC). The proposed Work could be carried around the
public key from the private key method supports large big data. boundary condition of the spline
Resistance to active collision and curve. Work can be extended to
replay attacks support digital signature standards
(DSS)
(Zolfaghari Study the cross-impact Detailed study on neural network No technique was proposed. The
and Koshiba, of neural network on and cryptography future where two data hiding
2022) cryptography techniques should be intersected
(Adnan and Enhancing secrecy, privacy, Affordable insights into 3D-AES cryptography.
Ariffin, 2019) confidentiality, and enhancing protection while Enhancement is needed to protect
availability to records from removing vulnerabilities cloud storage
attacks and threats while in
communication
282 Hybrid security of EMI using edge-based steganography and three-layered cryptography
(Adee and Securing and private data More redundancy, flexibility, RSA with AES then identity-based
Mouratidis, using cryptography with efficiency, and secrecy as it encryption algorithms followed
2022) steganography on cloud protects confidentiality, privacy, by LSB steganography. Future
environment leading to and integrity from attackers research work needs to focus on
reduced data theft and data while enhancing security and improving the combination of
manipulation attacks privacy steganography with cryptography
thus enhancing the security
(Maata, Information security to big Size of message is increased Two-Fish cryptography
Cordova, data in terms of size while significantly and time spent
and Halibas, transmitted efficiently and during the encryption and
2020) effectively decryption process. The authors
concluded that it was efficient
and effective
(El-Shafai et Securing images while in Secure, efficient, and immune SAE with improved deep learning
al., 2022) communication from various attacks such as (DL) extraction in the region of
noise attacks. This cryptosystem interest in the medical images
is efficient due parallelism of then compression and finally
the stacked auto-encoder (SAE), watermarking in multistage
which reduces the computational security encryption to enhance
complexity the robustness of medical data
broadcasted in telemedicine
(Akkasaligar To ensure and implement Resistance against different Selective digitizer medical image
and Biradar, security and confidentiality types of attacks. This SEDMI sncryption (SEDMI)
2020) of the medical images that method takes less computation
belongs to a larger data set or time (0.236 s) increasing its
larger size in bytes applicability as an e-health care
application
(Awadh, Image security and capacity Image quality is 68%, solving Hybrid layers of security
Alasady, and needs to be ensured on Internet security, and capacity concerns compression using discreet wavelet
Hamoud, transform (DWT) with AES
2022) encryption then least significant
bit (LSB) for hiding. Improved
hybrid security method with
random hiding algorithm to be
implemented on other languages
(Avula To develop a scalable, Developed a lightweight A Merkle tree data structure
Gopalakrishna lightweight framework based framework in blockchain. is used for hashing then
and Basarkod, on blockchain as modern Enhanced accessibility as cryptography based on lattice-
2023) healthcare are complex and artificial intelligence is combined based homomorphic proxy
requires secure storage with blockchain re-encryption scheme and
securely stored using blockchain
interplanetary file system
The role of cryptography is to ensure confidentiality The reverse of the proposed PHM method is applied
(Zolfaghari and Koshiba, 2022; Sharma and Prabha, to extract back the X-ray images.
2023) of EMI. The data set used for implementing
the proposed hybrid method (PHM) consist of 5856 Normalized cover image Lena
X-ray images in JPEG format which are all of vary- Firstly, the dimensions of Lena image are increased
ing sizes, dimensions, and belong to unique patients to 1080 × 1080. RONI region in Lena is the region
at unique time. The first step in PHM is to normal- other than Lena therefore the background (Hachaj,
ize the cover image Lena. Then X-ray image is hid- Koptyra, and Ogiela, 2021). Region of no interest is
den in normalized cover image Lena one at a time detected with the magic wand tool freely available
using edge-based steganography (EBS) resulting in a online at Pixlr ([Link] (https://
stego-image which is then applied with three layers [Link]/ n.d.). The background region of the cover
of cryptography this results in a crypto-stego-image. image Lena is inserted with randomly generated
This crypto-stego image can be saved either centrally black-and-white noise. Figure 37.4 depicts the process
or on a distributed database on cloud environment. of normalizing the cover image Lena and the output is
Applied Data Science and Smart Systems 283
referred to as the normalized cover image Lena. This Goldbaum, 2018) in JPEG format which are hidden
normalized cover image Lena hides one X-ray image one at a time, few of these are shown in Figure 37.5.
at a time using edge-based steganography (EBS) meth- One of the X-ray images is loaded from the data set
ods where the X-ray image is equally hidden across all of 5856 X-ray images. This X-ray image is hidden in
the edges of normalized Lena towards its background the edges of the normalized cover image Lena (around
for all the three red, green, and blue (RGB) compo- Lena in the noisy background) achieved earlier and
nents separately. shown in Figure 37.4. Then the stego-image is applied
with three layers of cryptography resulting in crypto-
Proposed hybrid steganography with layered cryptog- stego image. A detailed explanation of PHM method
raphy method proposed in this article has been mentioned in algo-
The input for PHM is one image at a time from a total rithm 1 in Table 37.2.
data set of 5856 X-ray images (Kermany, Zhang, and The output achieved after implementing PHM are
the images in the form of noisy signal as shown in
Figure 37.6. The decryption method is the reverse of
the encryption algorithm.
Figure 37.5 Few of the 5856 X-ray images that act as input for the PHM
284 Hybrid security of EMI using edge-based steganography and three-layered cryptography
Table 37.2 The stego-encryption algorithm for the PHM
Algorithm 1: Algorithm for implementing the proposed hybrid EBS steganography with three-layers of cryptography for
enhancing security of EMI
Data set: The cover image: Normalized cover image Lena; The secret image: the chest X-ray image, total number of
X-ray images: 5856, format of X-ray image is JPEG, Dimensions: each X-ray image varies in dimensions and properties.
Step 1: Study the normalized Lena image and find the edges around Lena in the region of no interest. Get each edge pixel
value as (col, row) for each edge position and store separately into array variable say col which stores the column pixel
values, while row variable array where the row pixel values are stored.
Step 2: Separate the normalized cover image Lena into three parts based on red, green, and blue (RGB) components and
save these three components into three arrays Ir, Ig, and Ib.
Step 3: Load one X-ray image from a data set of 5856 X-ray images and convert it into a 1-D array say A.
Step 4: Equally embed the elements of A in the edges of Lena across Ir, Ig, Ib components using the col and row pixel
value found earlier in Step 1.
Step 5: Check for remaining elements in A save it as rem variable, then
Find p the minimum value in the row array
If ( column_length > rem)
Hide the remaining pixel values in p=p-1 row
Else: n=mod (column_length, rem), embed remaining element across multiple rows from (p – 1) till p-n, while
ensuring that (p-n!=0)
Step 6: Step 5 performed edge-based steganography (EBS). This stego-image will now be applied with three layers of
cryptography. Firstly, create an initial permutation (IP) table which is of the same dimensions as one of the three RGB
component of stego-image I.
Step 7: The three RGB components of the stego-image are now stored into variables say Sr, Sg, and Sb.
Step 8: The IP table created in Step 6 will be used as a substitution table on Sr, Sg, and Sb. Where the values of Sr, Sg,
and Sb variables will substitute based on IP table. This step will result in the process of confusion.
Step 9: The Sr, Sg, Sb achieved after Step 8 will each be further divided into two half’s first is the left half say Srl, Sgl, Sbl
and the right half say Srr, Sgr, Sbr.
Step 10: Each half Srl, Sgl, Sbl the right half say Srr, Sgr, Sbr are individually applied with a circular right shift by 4
columns.
Step 11: Then XOR the right half with the left half which results in the new left half while the old left half will become
the new right half.
Step 12: Combine the new left half Srl, with the new right half Srr which will form the new red component similarly Sgl
with Sgr and Sbl with Sbr generating the green and blue components.
Step 13: Combine these RGB components to get the stego-crypto image.
Step 14: Store this stego-crypto EMI.
Step 15: Analyze the computational time for implementing the proposed hybrid method for future analysis.
Step 16: Stop.
Data set size comparison method uses a larger data set with a total size of 1.16
Table 37.3 is a comparative analysis table where the GB made up of 5856 X-ray images having unique
sizes of the data sets involved in the previously stud- dimensions while belonging to unique individual.
ied articles have been tabulated and compared with
the proposed hybrid method. This table discusses the Encryption time
programming language used by the researcher, the size The time taken to perform the PHM where the
of the data set, where the data set was downloaded Normalized cover image Lena is applied with
from, the type or format of the secret message, and its EBS Steganography then three-layer cryptography
sizes, similarly the file type and size of the cover image which generates a crypto-stego image. The encryp-
used with respect to the article studied previous stud- tion time also measures the encryption speed rate
ies in this article. and the throughput time therefore how fast the
On analysis, it was observed that the most popu- crypto-stego image will be generated. The encryp-
larly used programming language by researchers is tion time also helps determine whether the pro-
MATLAB which has also been used for PHM method posed method is suitable for real-time applications
implemented in this research work. Similarly, PHM or not. The proposed methods take a total time of
Applied Data Science and Smart Systems 285
Figure 37.6 Output images achieved after implementing the proposed PHM method
Table 37.3 Comparison based on the size of the data set and programming language used in previous research
Cited as Language Data set from / data set size Secret message type Cover type / cover size
/ secret message
size
PHM MATLAB Mendeley / 1.16 GB 5856 X-ray images Lena normalized the
in JPEG format / JPEG image / 720 KB
1.16 GB
(Ali et al., 2022) Python Log files 70% data training -/- -/ -
while 30% data for testing
purposes /-
(Lin et al., 2023) MATLAB 9.0 Head snapshots of 100 .JPEG images each -/-
version children (facial expression of 227 × 227 pixels
with ten EMI of hand X-ray)
(Mothi and MATLAB CASIA V4 and UBIRIS V1 Personal details of 10 iris images
Karthikeyan, 2019) 2017b Iris Databases. / 100 iris patients as text
images
(Georgieva-Tsaneva, MATLAB, -/- -/- records of up to 72 h of
Bogdanova, and Microsoft real electrocardiographs,
Gospodinova, 2022) Visual C++ photoplethysmography,
and Holter cardio data
(Kumar et al., 2022) - -/- Text message Digital image /-
“NATURE” /-
(Kore and Patil, 2022) Network -/- -/- -/-
simulator
(NS2)
(Adnan and Ariffin, - -/67240448 bits Big data /- -/-
2019)
(Adee and Mouratidis, Python Block of characters Text message 3 images/ 1.2MB, 2.9,
2022) converted to ASCII / - “Rose Adee and 7.2MB
encrypted files” /-
(Maata, Cordova, and Java [Link]/datasets/ Application store -/-
Halibas, 2020) 341.675 MB data /-
(El-Shafai et al., 2022) MATLAB -/ 256 × 256 - Grayscale images/ -
2020b and
Python
286 Hybrid security of EMI using edge-based steganography and three-layered cryptography
Cited as Language Data set from / data set size Secret message type Cover type / cover size
/ secret message
size
(Akkasaligar and MATLAB National Library of 500 medical 512 × 512 / -
Biradar, 2020) R2015b Medicine’s Open Access images (MRI,
Biomedical Images Search CT-Scan, X-Ray,
Engine /- and Ultra-Sound) /-
(Awadh, Alasady, and Visual Basic. -/- Lena image 65,536 Image 3,93,216 /-
Hamoud, 2022) Net language Bit /-
(Avula Gopalakrishna Ethereum [Link] -/- -/-
and Basarkod, 2023) platform and health./ 3452 health-related
Python data records of COVID-19
Table 37.4 Encryption time for the proposed hybrid Performance evaluation tests
method Performance evaluation test help prove whether the
aimed objectives have been achieved or not. Also,
Total time Average Minimum Maximum
time time time provide a better understanding of the performance
of any proposed hybrid technique (Li et al., 2022;
12.695647028333 0.13008 s 0.04451 s 0.548654 s Sharma and Prabha, 2023). These tests validate the
min
integrity, robustness, validity, and authenticity of the
proposed method (Ahmad et al., 2022). The values
achieved after implementing the proposed PHM
Table 37.5 Proposed method decryption time method are tabulated and compared with previ-
ously studied literature in Table 37.6. Some common
Total time Average Minimum Maximum performance evaluation tests such as MSE, RMSE,
time time time
etc. are discussed with their equations (Sharma and
16.369629188333 0.16772 s 0.05456 s 0.527336 s Prabha, 2023):
min
Mean square error (MSE)
MSE measures the average square error in image
retrieved after reversal of PHM when compared with
12.695647028333 min to perform the proposed
original image shown in Equation (1).
hybrid steganography with the encrypting method
(shown in Table 37.4). Thus, proving that the pro-
posed method is better than the previously stud-
(1)
ied methods having time 34.427972 (Adee and
Mouratidis, 2022), 102.164 s (Maata, Cordova,
and Halibas, 2020). Here, I(i,j) is the original image, SI(i,j) is the image
retrieved after implementing the reverse of the pro-
Decryption time posed hybrid technique. While m × n is the total num-
The time taken to get back the X-ray image hidden ber of pixels.
with PHM in the normalized Lena cover image. Lesser
decryption time is preferred. The proposed methods Peak signal to noise ratio (PSNR)
performed decryption in time of 0.05456 s as quoted The higher value of PSNR indicates that a low amount
in Table 37.5. Thus, achieving reduced computational of noise is present in the extracted image. MSE is rep-
time and is hence suitable for real-time applications. resented in Equation (2) (Sharma and Prabha, 2023).
Further from Table 37.5, it was found that the time
taken to perform encryption seems to be more, but it
(2)
is to be noted that the size of the data involved in this
study is also more than the previous studies. The aver-
age time for encryption is 0.13008 s while for decryp- Structural similarity index metrics (SSIM)
tion is 0.16772 s for 5856 X-ray images whose total SSIM measures the amount of similarity between
size is 1.16 GB. Thus, it can be concluded that the original image and retrieved image it is shown in
time of encryption and decryption has been reduced. Equation (3).
Table 37.6 Comparison based result table for the proposed hybrid method with the research article that led to this research
Cited as Time MSE / PSNR R SSIM NPCR UACI Entropy/ BER FSIM/ CR KLD/ RMSE/ SNR
MAPE PRD
PHM Total encryption time 0.000000035/ 0.999999 0.999981 95.585 6.59E-09 7.8398/2.25E- -/ 0.995818 0.000249/ 0.057113/ 0.049219
= 12.69 min, total 74.5584 05 1.68E-06 0.00011
decryption time =
16.36 min
(Al Hamid et 83.26 s - - - - - - - - - -
al., 2017)
(Lin et al., Avg. encryption and -/ 105.2513 0.9125 0.9406 100.00% 78.01% - - - - -
2023) decryption time of dB
0.065 s, 0.107 s.
(Georgieva- - 0.043/ 49.108 - - - - -/ 0.005 -/ 3.87 0.002/ 0.2074/ Ranges
Tsaneva, dB (PPG) to 0.0041 +- 0.164425 31.83–
Bogdanova, and 5.07 (Holter 0.001 (%) 46.37
Gospodinova,
2022)
(Kumar et al., - 0.0019922/ - - - - - - - - -
2022) 75.1375
(3) (10)
(6)
(14)
NPCR, UACI, PRD, and entropy. Hence proving that (2022). Hiding patients’ medical reports using an en-
the proposed PHM method ensures secrecy and pri- hanced wavelet steganography algorithm in DICOM
vacy of the X-ray images. images. Alexandria Engg. J., 61(12), 10577–10592.
[Link]
Akkasaligar, Prema T., and Sumangala Biradar. (2020).
Conclusion Selective medical image encryption using DNA cryp-
tography. Information Security Journal: A Global Per-
The PHM enhances the security of EMI while in stor-
spective. 29(2): 91–101.
age, or before being transmitted over network. The
Ali, Aitizaz, Muhammad Fermi Pasha, Jehad Ali, Ong
proposed hybrid method is implemented efficiently Huey Fang, Mehedi Masud, Anca Delia Jurcut, and
where a combination of edge-based steganography Mohammed A. Alzain. (2022). Deep learning based
with three-layered cryptography is implemented. The homomorphic secure search-able encryption for key-
X-ray image is firstly hidden with the help of edge- word search in blockchain healthcare system: A novel
based steganography and then applied with layered approach to cryptography. Sensors, 22(2): 528. 1–29.
cryptography which ensures secrecy, privacy, reduced Avula Gopalakrishna, Chandini, and Prabhugoud I. Basar-
computational time, and reduced computational kod. (2023). An efficient lightweight encryption model
cost thus making it suitable for real-time application with re-encryption scheme to create robust blockchain
and uses. The performance of the proposed hybrid architecture for COVID-19 data. Transactions on
Emerging Telecommunications Technologies. 34(1):
method is estimated by measuring the total amount
e4653. [Link]
of data thus 5856 X-ray images that are to be secured,
Awadh, W. A., Alasady, A. S., and Hamoud, A. K. (2022).
encryption time, decryption time, and total time. On Hybrid information security system via combination
comparative analysis with previously cited research, it of compression, cryptography, and image steganog-
was observed that the proposed method took a time raphy. Int. J. Elec. Comp. Engg., 12(6), 6574–6584.
of 0.13008 s for encryption and 0.16772 s for decryp- [Link]
tion which are lesser than the peers on comparison El-Shafai, W., Khallaf, F., El Sayed M. El-Rabaie, and Abd
with respect to size of the data set involved (here 1.2 El-Samie, F. E. (2022). Proposed neural SAE-based
GB). The PHM achieved a PSNR of 74.55 decibel medical image cryptography framework using deep
(dB) which is better, while MSE is close to zero which extracted features for smart IoT healthcare applica-
is preferred. With better values for SSIM, RMSE, tions. Neural Comput. Appl., 34(13), 10629–10653.
[Link]
MAPE, BER, etc. Thus, it can be concluded that the
Georgieva-Tsaneva, Galya, Galina Bogdanova, and Ev-
proposed hybrid method (PHM) is efficient and effec-
geniya Gospodinova. (2022). Mathematically Based
tive in securing a large data set of EMI images of Assessment of the Accuracy of Protection of Cardiac
varying dimensions and sizes. The statistical analysis Data Realized with the Help of Cryptography and
proves that PHM is better thus it enhances the secrecy Steganography. Mathematics, 10(3): 390. 1–18.
and privacy of EMI. In the future, a machine learning Brar, P. S., Shah, B., Singh, J., Ali, F., and Kwak, D. (2022).
algorithm can be implemented for easy detection of Using modified technology acceptance model to eval-
edges in the cover image, and the cover image Lena uate the adoption of a proposed IoT-based indoor
can be changed to any image in general. Further, it can disaster management software tool by rescue work-
be integrated into the blockchain environment. ers. Sensors, 22(5), 1866, [Link]
s22051866.
Hachaj, T., Koptyra, K., and Ogiela, M. R. (2021). Eigenfac-
References es-based steganography. Entropy, 23(3), 1–24. https://
Adee, Rose, and Haralambos Mouratidis. (2022). A dy- [Link]/10.3390/e23030273.
namic four-step data security model for data in cloud Hamid, H. A. A., Mizanur Rahman, Sk Md, Hossain, M. S., Al-
computing based on cryptography and steganography. mogren, A., and Alamri, A. (2017). A security model for
Sensors, 22(3): 1109. 1–23. preserving the privacy of medical Big Data in a health-
Adnan, N. A. N. and Ariffin, S. (2019). Big data security care cloud using a fog computing facility with pair-
in the web-based cloud storage system using 3d- ing-based cryptography. IEEE Acc., 5, 22313–22328.
Aes block cipher cryptography algorithm. Comm. [Link]
Comp. Inform. Sci., 937, 309–321. [Link] Https://[Link]/. (n.d.). Accessed March 18, 2022. https://
org/10.1007/978-981-13-3441-2_24. [Link]/.
Agarwal, Shweta, and Chander Prabha. (2022). Analysis Kermany, Daniel, Kang Zhang, and Michael Goldbaum.
of Lung Cancer Prediction at an Early Stage: A Sys- (2018). Labeled optical coherence tomography (oct)
tematic Review. In Congress on Intelligent Systems: and chest x-ray images for classification. Mendeley
Proceedings of CIS 2021. 1, 701–711. Singapore: data. 2(2): 651.
Springer Nature Singapore. Kore, A. and Patil, S. (2022). Cross layered cryptography
Ahmad, M. A., Elloumi, M., Samak, A. H., Al-Sharafi, A. based secure routing for IoT-enabled smart health-
M., Alqazzaz, A., Kaid, M. A., and Iliopoulos, C. care system. Wire. Netw., 28(1), 287–301. [Link]
org/10.1007/s11276-021-02850-5.
290 Hybrid security of EMI using edge-based steganography and three-layered cryptography
Krishna, A. V. N. (2018). A Big–Data security mechanism integrity for intelligent application. Int. J. Elec. Comp.
based on fully homomorphic encryption using cubic Engg., 13(4), 4422–4431. [Link]
spline curve public key cryptography. J. Inform. Op- ijece.v13i4.pp4422-4431.
tim. Sci., 39(6), 1387–1399. [Link] Prabha, C., Singh, J., Agarwal, S., Verma, A., and Sharma,
2522667.2018.1507762. N. (2022). Introduction to computational intelligence
Kumar, M., Soni, A., Shekhawat, A. R. S., and Rawat, A. in healthcare. Computat. Intel. Healthcare, 1–15.
(2022). Enhanced digital image and text data security [Link]
using hybrid model of LSB steganography and AES Sharma, D. and Kawatra, R. (2023). Security techniques
cryptography technique. Proc. 2nd Int. Conf. Artif. implementation on big data using steganography and
Intel. Smart Energy, ICAIS 2022, 1453–1457. https:// cryptography. Lec. Notes Netw. Sys., 517, 279–302.
[Link]/10.1109/ICAIS53314.2022.9742942. [Link]
Li, C., Dong, M., Li, J., Xu, G., Chen, X. B., Liu, W., and Sharma, D. and Prabha, C. (2023). Security and pri-
Ota, K. (2022). Efficient medical big data manage- vacy aspects of electronic health records: A review.
ment with keyword-Searchable encryption in health- 2023 Int. Conf. Adv. Comput. Comp. Technol. (In-
chain. IEEE Sys. J., 16(4), 5521–5532. [Link] CACCT), 815–820. [Link]
org/10.1109/JSYST.2022.3173538. CACCT57535.2023.10141814.
Lin, C. H., Wen, C. H., Lai, H. Y., Huang, P. T. , Chen, P. Y., Sharma, N. and Prabha, C. (2021). Computing paradigms:
Li, C. M., and Pai, N. S. (2023). Multilayer convolu- An overview. 2021 Asian Conf. Innov. Technol.
tional processing network based cryptography mecha- (ASIANCON), 1–6. [Link]
nism for digital images infosecurity. Processes, 11(5). CON51346.2021.9545007.
[Link] Sharma, V., Singh, T., Garg, N, Dhiman, S., Gupta, S.,
Maata, R. L. R., Cordova, R. S., and Halibas, A. (2020). Rahman, Md., Najda, A., et al. (2021). Dysbiosis
Performance analysis of twofish cryptography algo- and Alzheimer’s disease: A role for chronic stress?
rithm in big data. ACM Int. Conf. Proc. Ser., 56–60. Biomolecules, 11(5), 678. [Link]
[Link] biom11050678.
Singh, J., Goyal, G., and Gill, R. (2020). Use of neuro- Shukla, P. K., Sandhu, J. K., Ahirwar, A., Ghai, D., Ma-
metrics to choose optimal advertisement method for heshwary, P., and Shukla, P. K. (2021). Multiobjec-
omnichannel business. Enterp. Inform. Sys., 14(2), tive genetic algorithm and convolutional neural
243–265, [Link] network based COVID-19 identification in chest
40392. X-ray images. Math. Prob. Engg. 1–9. [Link]
Mothi, R. and Karthikeyan, M. (2019). Protection of bio org/10.1155/2021/7804540.
medical iris image using watermarking and cryp- Zolfaghari, Behrouz, and Takeshi Koshiba. (2022). The
tography with WPT. Meas. J. Int. Meas. Confeder., dichotomy of neural networks and cryptography:
136, 67–73. [Link] War and peace. Applied System Innovation. 5, no. 4
ment.2018.12.030. (2022), 5, 1–28.
Parmar, M. and Shah, P. (2023). Internet of things-Block-
chain lightweight cryptography to data security and
38 Efficient lung cancer detection in CT scans through
GLCM analysis and hybrid classification
Shazia Shamas, Surya Narayan Panda and Ishu Sharmaa
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
Timely detection of lung cancer is important, significantly impacting patient prognosis and decreasing mortality rates. com-
puted tomography (CT) scans have become a cornerstone in this endeavor due to their ability to provide detailed anatomical
information. However, a persistent challenge in this field is striking the delicate balance between precision accuracy, and
execution time during the detection process. Existing precision-focused methods often demand extensive computational
resources, leading to prolonged execution times – undesirable in time-sensitive clinical scenarios. This paper introduces a
groundbreaking solution by proposing a novel hybrid classification algorithm for CT image analysis. The algorithm achieves
exceptional precision while substantially reducing execution times. It integrates gray-level co-occurrence matrix (GLCM)
analysis into its core, efficiently identifying cancerous regions within CT scans. This approach comprises a sequential pro-
cess: GLCM analysis, feature extraction, hybrid classification, algorithm training, and detection, resulting in high-precision
and accurate lung cancer detection within minimal execution time. From the results, it is clear that SURF surpasses SIFT
with a minimum error rate of 16.71 compared to SIFT’s 39.02. SURF also executes faster, taking 0.096 s vs. SIFT’s 3.46 s.
As a result, SURF is expected to have superior recall and precision. Hence, this research addresses a critical need in the field,
offering a promising pathway toward expedited, precise, and scalable lung cancer diagnosis.
Keywords: Lung cancer, timely detection, CT scans, gray-level co-occurrence matrix (GLCM) analysis, feature extraction,
hybrid classification
a
[Link]@[Link]
292 Efficient lung cancer detection in CT scans through GLCM analysis and hybrid classification
a novel hybrid classification algorithm tailored for employing bag of visual words (BOVW) based on
computed tomography (CT) image analysis. The algo- K-means clustering for attributes extracted using
rithm endeavors to achieve exceptional precision in SIFT in the preliminary stage. Subsequently, a super-
lung cancer detection while significantly reducing exe- vised learning algorithm, BPNN, a subset of artificial
cution times. A pivotal breakthrough lies in the seam- neural networks (ANN), was employed for classifica-
less integration of gray-level co-occurrence matrix tion. Finally, the watershed segmentation method was
(GLCM) analysis into its core, facilitating efficient utilized to detect nodules in cancerous lung images.
identification of cancerous regions within CT scans. The validation results demonstrated an impressive
So, to address the challenging issue of identify- accuracy of 91% for this established technique when
ing and classifying the cancerous areas in the scans compared to various other algorithms (Basha et al.,
efficiently in relation to high precision and mini- 2020).
mum execution time, this research article proposes This paper introduced a study where medical images
a novel approach of hybrid classification algorithms underwent analysis using image processing, machine
integrated with the GLCM. The proposed integrated learning (ML), and complementary technologies to
approach follows a sequence of steps in order to solve detect and address cancer at its early stages within
the challenging issue highlighted in this paper. The contemporary clinical settings. Their proposal cen-
sequential process includes GLCM analysis, feature tered on an automated approach utilizing CT images
extraction, hybrid classification, algorithm train- to identify lung cancer in its nascent phase, aiming
ing, and detection, which results in high-precision for a high standard of performance accuracy. A novel
and accurate lung cancer detection within minimal framework was devised for diagnosing lung cancer,
execution time. This study offers an achievable path involving extraction of various attributes from CT
towards quick, accurate, and scalable detection, suc- scans and subsequent stages such as image enhance-
cessfully filling a major gap in the area of lung cancer ment, segmentation, feature extraction, and applica-
diagnosis. tion of a support vector machine (SVM). Ultimately,
experimental results demonstrated the superior accu-
racy of their recommended technique (Hoque et al.,
Related work
2020).
This section contains a thorough analysis of the per- The aim of this paper was to direct their efforts
tinent literature. A variety of algorithms have been toward devising a system to detect lung cancer utiliz-
investigated by numerous researchers with the goal of ing CT scan images. The system involved four integral
identifying lung cancer. The level of exploration and phases. Initially, CT scan images were pre-processed to
research into these algorithms, meanwhile, has been enhance image quality. Subsequently, the anticipated
rather constrained. cancerous object was identified and isolated from the
The authors in this article used deep learning background through segmentation. Features, such as
models based on artificial intelligence (AI) for auto- area and energy, were extracted from the identified
matically detecting malignant cells in the lungs. The objects. This allowed for the classification of lung
examination analyzed the performance of four diverse cancer into cancerous and non-cancerous categories.
AI frameworks for detecting lung nodule cancer such The system they presented exhibited a precision of
that the doctors/radiologists could provide accurate 83.33% in effectively detecting lung cancer (Firdaus
diagnostic results. The two experienced doctors with et al., 2020).
more than 10 years of involvement in the fields of The authors in this paper proposed a method for
aspiratory basic consideration, and emergency clinic detecting lung cancer from chest CT images using
medication selected a sum of 648 samples. A number co-learning and clinical demographics. Over the last
of metrics (e.g., curve receiver operating characteristic decade, image-processing methods have gained sig-
curve (ROC), area under the curve (AUC), accuracy, nificant traction across various clinical domains for
specificity, etc.) were considered in this work for mea- cancer detection and treatment. Time played a crucial
suring and evaluating the results generated by the pre- role in identifying anomalies in input images. Swift
sented model. This hybrid deep neural network was detection of diseases relied on accuracy and image
best in class design, with superior accuracy and low quality, emphasizing the importance of image quality
FP outcomes. Doctors use this automatic framework evaluation during enhancement stages. Various image
to safeguard a quality relationship between doctors processing techniques, including image enhancement,
and patients (Nadkarni et al., 2019). segmentation, and feature extraction, have proven
This research paper developed an automated effective in detecting tumors within images. The
lung cancer detection system utilizing a combina- development of a computer-aided diagnosis (CAD)
tion of SIFT, enhanced wavelet transforms, BPNN, system for lung cancer detection was rooted in an
and watershed segmentation. The process involved integrated approach combining image processing and
Applied Data Science and Smart Systems 293
ML methodologies. Extending beyond image process- with 88.8% sensitivity. This work considered a num-
ing, lung cancer diagnosis involved feature extraction ber of features to model the nodule growth predic-
and selection following segmentation. The proposed tion measure. The overlay of these events for larger,
approach effectively identified cancerous cells from average, and minimal nodule growth cases was not as
CT scans, positioning lung CT scans as the primary much. Hence, it was possible to use this constructed
data source in this innovative strategy (Pranathi et al., growth prediction model to help doctors while mak-
2019). ing decisions on the malignant nature of lung nod-
This paper focused on improving accuracy and ules from a previous CT image (Krishnamurthy et al.,
precision in the early-stage detection of lung cancer. 2017).
To achieve this, they integrated biomedical image This paper suggested a GLCM model in order
processing methods with knowledge discovery in to extract the lung images of patients. This model
databases (KDD). The lung images from CT scan assisted in extracting 3 properties from the grow-
data were subject to pre-processing and segmenta- ing ROI. The levels of lung cancer were detected by
tion in the region of interest (ROI). Subsequently, computing the nodules with these properties. Diverse
various attributes were categorized using the random levels of the tumor were represented through the size
forest (RF) technique, leveraging the SURF algo- of the nodule. The SVM algorithm was applied to
rithm. SVM algorithm was then utilized for feature detect the abnormal lung image. The generalization
extraction. The classification process determined controls were put together with a strategy so that the
whether the image depicted a healthy or unhealthy dimension of nodules was addressed. The margin was
state. The technique’s performance was evaluated increased which had consistency with the weights for
using a function evaluation plot, employing both RF obtaining the generalization control during the clas-
and SVM. Remarkably, the SVM yielded the most sification issues (Jony et al., 2019).
favorable results. The process achieved an efficiency The authors in this study recommended an effective
rating of 94.5%, with a sensitivity of 74.2% and a algorithm to detect and predict lung cancer in which
specificity of 77.6% (Kyamelia et al., 2019; Gill et the SVM classification algorithm was deployed. The
al., 2020). cancer was detected by applying the multi-stage clas-
This paper endeavored to integrate AI into the sification. In each phase, the image was enhanced
medical domain, specifically to diagnose diseases in and segmented. Different processes were carried to
their early stages. Their approach involved process- perform the image enhancement. The image was seg-
ing images using CT scans as input data sourced mented through 16 pages – the threshold and marker-
from the lung image database consortium (LIDC). controlled watershed-based segmentation. The SVM
The initial pre-processing phase entailed converting was implemented to execute the classification process.
RGB images into grayscale and subsequently into It was analyzed that the recommended algorithm pro-
binary images. After that, the binary images were fed vided superior precision while detecting lung cancer
to the convolution neural network (CNN) for detect- (Alam et al., 2018).
ing and side-by-side classifying the images as can- This paper presented the RBFNN classification
cerous or noncancerous. During the whole analysis technique for detecting whether the lung was affected
performed by the system. The major contribution of by cancer or not. The GLCM technique was deployed
this article was to design a system that can classify with the objective of extracting attributes from the
these images with minimum utilization of power and chest radiograph. These attributes were computed in
time in order to enhance lung cancer (Rohit et al., order to carry out the detection procedure of the pre-
2019). sented technique. There were 5 attributes comprised
This paper aimed to recognize the cancerous lung for this purpose. The outcomes revealed that the
nodules accurately and at an early stage with fewer image enhancement were efficient in the maximiza-
FPs (false positives). This paper segmented all possi- tion of the accuracy of the presented technique for
ble nodule candidates using auto center seed k-means detecting lung cancer with the help of the chest radio-
clustering algorithm based on block histogram. This graph (Miah et al., 2015).
work computed effective shape and texture features
(2D and 3D) for eliminating untrue nodule candidates. Objectives
This work performed the classification of cancerous
and non-cancerous tumors using a two-stage classi- The main aim of this research article is to identify the
fication model. The initial phase using a rule-based lung cancer regions, particularly at early stages. This
classifier produced a sensitivity of 100% but with a article mainly focuses on balancing the challenging
high false positive of 13.1 for every patient image. In issues of execution time, precision, and accuracy dur-
the next phase, a BPN-based ANN classifier was uti- ing the detection process. The main challenging objec-
lized to reduce the false positive to 2.26 for each scan tives are briefly discussed below.
294 Efficient lung cancer detection in CT scans through GLCM analysis and hybrid classification
Conclusion
To summarize, timely identification of lung cancer is
vital for enhancing patient prognosis and minimiz-
ing mortality rates. CT scans have revolutionized this
process by providing intricate anatomical details, yet
a delicate balance between precision and execution
time remains a challenge. Current precision-focused
methods often demand extensive computational
resources, resulting in undesirable delays in critical
clinical scenarios. This study presents an innovative
hybrid classification algorithm for CT image analysis,
revolutionizing lung cancer detection. By integrating
GLCM analysis, this algorithm efficiently pinpoints
cancerous regions within CT scans, ensuring excep-
tional precision and significantly reduced execution
times. The sequential process – GLCM analysis, fea-
ture extraction, hybrid classification, algorithm train-
ing, and detection – demonstrates high-precision
and accurate lung cancer detection within minimal
Figure 38.4 SIFT feature extraction
execution time. The comparison of SURF and SIFT
highlights the superiority of SURF in terms of error
rate and execution speed, signifying its potential for
enhanced recall and precision. Consequently, this
research addresses a critical gap in the field, provid-
ing a promising avenue toward rapid, precise, and
scalable lung cancer diagnosis. The novel approach
proposed here has the potential to reshape lung can-
cer detection, bringing us closer to more effective and
timely medical interventions.
References
Nadkarni, S. and Borkar, S. (2019). Detection of lung cancer
in CT images using image processing. Proc. Int. Conf.
Trends Elec. Informat. (ICOEI), 57–65.
Zeelan Basha, C., Lakshmi, B., Vineela, D., and Lakshmi, S.
(2020). An effective and robust cancer detection in the
lungs with Back Propagation Neural Networks and
watershed segmentation. Int. J. Recent Technol. Engg.,
8(3), 200–220.
Figure 38.5 SURF feature extraction
Hoque, A., Farabi, A., Fahad, A., and Zahid, M. (2020).
Automated detection of lung cancer using CT scan im-
ages. Proc. Int. Symp. Comp. Sci. Intel. Con. (ISCSIC),
46–53.
Firdaus, Q., Sigit, R., Harsono, T., and Anwar., A. (2020).
Lung cancer detection based on CT-scan images with
detection features using gray level co-occurrence
matrix (GLCM) and support vector machine (SVM)
methods. Proc. Int. Elec. Symp. (IES), 212–219.
Pranathi, K., Suvarna Vani, K., Praveen, K., and Koduru, J.
(2019). Lung cancer detection using CT Scan image.
Adv. Computat. Bio-Engg., 1(1), 233–243.
Kyamelia, R., Sinha, C., Madhurima, B., Ganguly, A., Dutta,
C., and Banik, R. (2019). A comparative study of lung
cancer detection using supervised neural network.
Proc. Int. Conf. Opt-Elec. Appl. Optics (Optronix),
Figure 38.6 Graphical comparison between SIFT and 211–218.
SURF
Applied Data Science and Smart Systems 297
Rohit, Y. B., Harsh, P. J., Rachana, K., Gaitonde and Raut, Jony, M., Tujohora, F., and Rana, H. (2019). Detection of
G. (2019). A novel approach for detection of lung lung cancer from CT scan images using gray scale co-
cancer using digital image processing and convolu- occurrence matrix and support vector machine. Proc.
tion neural networks. Proc. Int. Conf. Adv. Comput. Int. Conf. Adv. Sci. Engg. Robot. Technol. (ICASERT),
Comm. Sys. (ICACCS), 223–230. 71–83.
Krishnamurthy, S., Narasimhan, G., and Rengasamy, U. Alam, J. and Hossan, A. (2018). Multi-stage lung cancer
(2017). An automatic computerized model for cancer- detection and prediction using multi-class SVM classi-
ous lung nodule detection from computed tomogra- fier. Proc. Int. Conf. Comp. Comm. Chem. Mat. Elec.
phy images with reduced false positives. Rec. Trends Engg. (IC4ME2). 35–42.
Image Proc. Pat. Recogn., 343–355. Miah, M. B. A. and Yousuf, M. A. (2015). Detection of lung
Gill, R. and Singh, J. (2020). A review of neuromarketing cancer from CT image using image processing and
techniques and emotion analysis classifiers for visual- neural network. 2015 Int. Conf. Elec. Engg. Inform.
emotion mining. 2020 9th Int. Conf. Sys. Model. Adv. Comm. Technol. (ICEEICT), 1–6. doi: 10.1109/ICEE-
Res. Trends, (SMART), 103–108. ICT.2015.7307530.
39 Newton Raphson method for root convergence of higher
degree polynomials using big number libraries
Taniya Hasija1, K. R. Ramkumar2,a, Bhupendra Singh3, Amanpreet Kaur4
and Sudesh Kumar Mittal5
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
1,2,4,5
3
Centre for Artificial Intelligence & Robotics, Defence Research and Development Organization, Bangalore, India
Abstract
Polynomial root discovery is applicable to cryptography domain in many aspects. There are number of methods such as
bisection, Newton Raphson, and Secant being used to discover a possible root of a random polynomial. However primitive
data types available with compilers are limiting the root convergence to the 15th degree of any polynomial effectively. In
cryptography, polynomials with higher degrees can increase confidentiality levels and make a sustainable key against attacks
from both classical and quantum computers. This paper reveals a method of using big number libraries for converging a root
of a given higher degree polynomial, with proper verification, this can be applied to post quantum cryptographic algorithms
for encryption and decryption.
Keywords: Newton Raphson method, root finding algorithm, big number in C, polynomials, enterprises, security
[Link]@[Link]
a
Applied Data Science and Smart Systems 299
(Lang and Frenzel, 1994). Hansen and Patrick have Newton Raphson method
evaluated a number of iterative strategies for finding
A solid approach for numerically fathoming equa-
the nonlinear equation’s root. They included Halley,
tions is the Newton-Raphson method. A real-valued
Euler, Ostrowski, Lagurree, and Newton techniques
function with the root f(x) = 0 can be easily approxi-
in their analysis part of algorithms. Newton is a
mated using the Newton-Raphson method (Akram
quadratic convergent, but Lugerree, Halley, Euler,
and Ann, 2015). The Newton-Raphson algorithm
and Ostrowski are cubic convergent to the root. The
is predicated on the notion that approximation is
Laguerre technique is superior to other approaches
achieved by digression, which essentially involves
when the starting point is regarded as z for which |z|
computing the x-intercept of the digressing line,
is large (Hansen and Patrick, 1976). The fourth-order
starting with a prior assumption that is logically
convergent to root approach, developed from the
close to the root. It employs the continuous and
Newton Raphson method, was introduced by Chun
differentiable function to get the x-intercept (Ben-
(2006) in 2006. It does not require the second-order
Israel, 1966; Ypma, 1995). The equation of the
derivative of a function. The Adomian decomposi-
Newton method is developed from the slope of a
tion method (Adomian and Rach, 1985) has been
line.
modified to create this iteration. Darvishi and Barati
(2007b) proposed a novel, better Newton approach
Derivative of the Newton method
in which they converge to a root by third order
i. f(x) = 0 is a given equation
or cubic. They expanded Chun’s method in their
ii. Starting from an initial point x0
approach (Chun, 2006). Jacobian matrix at position
iii. Determine the slope of f(x) at x = x0. Termed it as
xn is used in the iterative method for solving non-
f'(x0)
linear equations. A further publication by Darvishi
and Barati (2007a) on the fourth-order convergence
of their equation from (2007b) and quadrature for-
mulae was released. In order to solve non-linear
equations, Noor and Waseem proposed a two-step
iterative method and demonstrated the cubic conver-
gence of their algorithms (Noor and Waseem, 2009).
The Newton method, Cordero and Torregrosa’s pro-
posed method (2007), and the method proposed by Algorithm of the Newton Raphson method
Darvishi and Barati (2007a) are also used to com-
pare these introduced methods. Sharma and Guha Input : coeff_arr: array that contains coefficients of the
polynomia function f(x), coeff_arr ∈ I
(2013) leveraged Homeier’s third-order convergence Output: root: root of the given polynomial
method to construct a three-step iterative method
Step 1: Choose an initial guess x0, let [a, b] be any
that is fifth-order convergence to root. A comparison
of various methods for finding roots of polynomials interval such that f(a)<0 and f(b)>0, then
is made by Chun et al. (2017) who conducted their Step 2: Set i=0 and Repeat step 3–5, until xi+1==xi
comparative analysis research on root finding up to Step 3: Calculate f(xi) and f'(xi) symbolically using
the convergence of the eighth order. They came to coeff_arr
the conclusion that the third algorithm provided by Step 4: Set
Dong is the best among all the algorithms mentioned Step 5: Increment i by 1.
in their paper. Neta et al. (2012) analyzed Halley and Step 6: xi required root of the polynomial tactically it is
Jarratt’s method for third and fourth order conver- a cipher text of a given polynomial.
Step 7: return root.
gence of roots, and it works well with non-linear
systems of equations using higher order iteration.
The aforementioned review makes it evidential that a Consider a polynomial as given in Equation (2)
variety of iterative techniques are used to find roots with an assumption that this polynomial is generated
of random polynomials, but very limited research is from a seed-polynomial.
done to determine the root of a higher degree poly-
nomial more than 100 degree that has big co-efficient (2)
values and constants.
This work implemented a specific version of An initial xi value is calculated from the nth root
root-convergence method with the help of big num- of the constant value, where n is the highest degree of
ber libraries of “C” language to check the suitabil- the polynomial. Here in Equation (2), highest degree
ity of Newton-Raphson method for cryptographic is 2, so square root is taken according to the given
applications. example. Computed value of x = = 3.46, taken
300 Newton Raphson method for root convergence of higher degree polynomials
as the initial x value and substituted in the Newton- additional computation (Kaushal, Bhardwaj et al.,
Raphson formula to compute the next xi values. This 2022).
procedure is repeated until two successive iterations
have the same computed x value. This x value is an GNU multiple precision arithmetic library (GMP)
approximate real root of a given polynomial func- Gnu’s not Unix (GNU) is an extensive collection of
tion. Derivative f'(x) = 4x + 2 is required to evaluate software which is free and can be used for software
Newton-Raphson method. The evaluation steps are purposes and can also be used as an operating sys-
given in Table 39.1. tem or part of an operating system (Stallman, 1985).
It is always suggested to take odd degreed polyno- GNU provides a set of libraries and packages that
mials to get accurate root values. Here we consider can be used for different areas. To deal with large
the first root of the polynomial. numbers and high precision values, GNU multiple
precision arithmetic library (GMP) is used. This
Big numbers in C language specific library can do arbitrary-precision arithmetic on large
integers, large rational numbers, and large floating
In C language, to deal with numbers and calcu- point values. There are no restrictions on variable
lations, integer and double data types are used. precision other than those imposed by available
The Integer ranges from -2,147,483,648 to memory (operands may be of up to 232−1 bits on
2,147,483,647 (32 bits), and 64 bits are allocated 32-bit machines and 237 bits on 64-bit machines)
for double data type, where 1 bit is for sign storage, (Granlund, 2015; Kaushal, Kumar et al., 2022).
exponent utilizes the 11 bits and the rest 52 bits There are a number of functions in the GMP library
are for storing mantissa. Meanwhile, a double data that deal with the arithmetic operation of two big
can support 15 decimal digits precision. A big num- number operands. “gmp.h” header file is included in
ber implementation in cryptography to manage big the C program to use the data types and functions of
sized keys, plain texts and cipher texts are found that library.
to be useful for better results (Singh et al., 2009;
Fujdiak et al., 2017). The public key cryptography
Experimental setup and implementation
RSA algorithm uses big prime numbers for encryp-
tion and decryption (Sarma and Avadhani, 2011). The programming is done in the C programming lan-
The usage of big numbers in private key cryptog- guage, and the system environment is Linux. GNU
raphy techniques like data encryption standards compiler collection (GCC) is a collection of compilers
(DES) and advanced encryption standard (AES) can that can compile a variety of programming languages
improve the overall efficiency and speed of encryp- such as C, C++, Fortran, and D. GCC is used to com-
tion and decryption. The string data type is used pile the C code in research work. To deal with the big
to store large number but retrieving number from numbers GMP header file is installed. This research
strings and doing arithmetic operations required work is able to handle big and complex calculations
1 2.26914414
2 2.013079343
3 2.000034036
4 2.000000000
5 2.000000000
Applied Data Science and Smart Systems 301
Figure 39.1 Flow chart of the procedure followed to implement 3 experiments using primitive and big number data
types and their accuracy and correctness evaluation
of polynomial equations using the GMP library. In tested with a number of polynomials in which orders
this work, three types of implementations have been are ranging from 1 to 15th degree, the results have
done for executing Newton Raphson code in C lan- become unstable after 15th degree polynomials
guage on the basis of primitive data types and big for 15-digit plain texts. The evaluation of the root
number libraries supported data types. The encapsu- is done by computing the f(x) function. If f(x) = 0,
lated flow chart of three experiments has been shown then root is correct else not. Some examples of this
in Figure 39.1. implementation are shown in Table 39.2. The root
First, the implementation of Newton Raphson convergence becomes unstable after 15th degree
code is done using primitive data types. We have polynomials, an encrypted data should be decrypted
302 Newton Raphson method for root convergence of higher degree polynomials
Table 39.2 Polynomial root finding with primitive data types
S. Degree Coefficients of Constant value of Computed root using Execution Evaluation Correct root
No. of the the polynomial the polynomial Newton Raphson time in of f(x) convergence
polynomial (double data (double data type) method (double data seconds
type) type)
with 100% accuracy, means that, an encrypted data As the polynomials are used in cryptography
should be always decrypted correctly. In Table 39.2, and other applications, there is a need of big num-
the 17th degree polynomial does not give accurate ber calculations, so that the accurate root can be
result and same is applicable to higher degree poly- generated from the polynomials that also have big
nomials. There is a need of better implementation numbers as their constant and coefficients, also can
options to encrypt big plain texts. deal with higher degree polynomials. In the second
Table 39.3 Polynomial with primitive co-efficient and big number constant
S. Degree Coefficients of the polynomial Constant value of Computed root using Execution Evaluation Correct
No. of the the polynomial Newton Raphson time in of f(x) root
polynomials method seconds convergence
1 101 -9222.93, 5150.62, -2596.19, 7926.57, 796.74, -9259.86, (768 bits) 77.736871368624825 0.024466 0.0 Yes
9335.39, 49.66, 995.76, -7415.86, 3726.71, -2141.33, 8777.05, 155251809230070 32613799396671043
9190.49, 2325.45, 1087.92, 7541.71, -5877.16, -7493.36, 893514897948846 98585333905694111
297.80, -9956.48, 8592.79, 429.41, -916.13, 7949.02, 250255525688601 60516019818824138
-7592.98, 8335.04, -1838.14, -5919.29, 1052.35, 7921.04, 71166966111350 93068595565209857
-9793.13, -8378.80, 5600.65, -9474.03, 4162.26, 5456.12, 947534894624925 42351689757797285
-8722.37, 1397.50, 7310.71, 7970.48, 1022.24, -4638.00, 52179522115341 54053227563484907
3417.62, -3889.01, 8802.04, 6610.23, -7272.80, -4265.46, 344282218821823 00279100288623609
-861.19, 8771.91, 3169.98, 7778.50, 113.56, -1530.85, 32895145639727 739141441848700936
8522.52, -2327.96, 5806.86, 2405.13, 7523.25, 3694.44, 898684564781135 16439331262189840
-8104.19, -4920.45, 5293.70, -8250.44, 2925.39, 1178.34, 911425012413524 80108357819023210
1250.37, -7987.73, 8132.84, 8788.21, -9019.92, 6200.11, 710743706422593 837975184369276576
8082.13, -4051.25, 4899.39, 1930.67, -8050.36, 8380.56, 331719098628765 09456495428313647
-6037.04, 9498.94, 3653.34, -5057.80, 9634.12, 6267.65, 602117654288074 27523761279819596
2564.11, -4422.31, 6172.03, 6410.81, -9227.67, 9859.54, 972916187976263 78691614788606072
-8207.75, -7130.63, 8641.33, -9602.16, -2142.18, -3640.80, 849004251546271 58153294410258164
-9778.38, -699.24, 1744.13, 7179.45 9946096640 463566440473431873
629255170043125870
2 151 -6618.81, 4225.80, 3445.51, 3336.85, -3501.72, -7981.85, (768 bits) 1552518 32.049606065605774 0.037989 0.0 Yes
1491.79, -4896.58, 4698.45, 6381.15, -2560.78, 123.55, 2965.98, 092300708848966 21637208919790806
8109.73, 2536.80, -2612.23, -8386.38, 5317.62, 74.42, -3066.43, 912877493951012 49097052695544485
5674.98, -2966.03, -8452.16, 2148.18, -5505.47, 3431.26, 620507776008668 00943802319409077
experiment, the degree of the polynomial is given big number libraries ranging from 1 to 1024 bits.
as input, after getting the degree of the polyno- After generating a complete polynomial, Newton
mial, the next task is to generate the polynomial. Raphson method is applied to compute a root.
For generating polynomial random number genera- The accurate root evaluation takes place till 115
tor is used. Here using a big number library, 64-bit degrees, afterward the polynomial gets the errone-
numbers are generated and served as the coefficient ous root. A few tested polynomials are shown in
of the polynomial, and a constant is generated Table 39.4.
which a big number is having lengths from 1 to In Table 39.4, the entries of 5th degree and 15th
1024 bits that can be varied according to our bit degree have shown even though it is compatible with
length choice. After that Newton Raphson code is the 115-degree polynomial. It is additionally seen
implemented and the root is computed. Then the that the most elevated coefficient and constant length
verification is done on the basis of f(x) = 0. Some is 1024 bit, therefore to store a precise root, 1024-
examples of the implemented work are shown in bit length is required for the root and the accuracy
Table 39.3. of the root is not depending upon the degree of a
Table 39.3 creates strong evidence that we can polynomial.
find the roots of polynomials with large constant Moreover, it has been seen that the root computa-
values and of any degree or order. The root conver- tion is exceptionally fast using big number libraries. The
gence happens till 201th degree polynomial success- time of root computation is given in Tables 39.2–39.4.
fully, means, encryption and decryption can be done As the degree of the polynomial expanded the time of
more accurately. The key length will be more than computation is increased but still it is milliseconds (ms)
2048 bits which is far better than AES algorithm range only. This implementation uses all big numbers
that uses three different key lengths (128, 192, and still it suffers after 115th degree polynomial as com-
256) with better memory utilizations. Figure 39.2 pared to previous one that works well for 201st degree
gives the memory requirement in bits details to sat- polynomials because it takes big values as co-efficient
isfy the f(x) = 0 test, where the degree of the polyno- values.
mial varies from 3 to 201. It is clear from the Figure Figure 39.3 depicts the temporal complexity graph
39.2 that in spite of any degree of the polynomial for the codes of experiments 1, 2, and 3. The graph
(from 5 to 201 degree) the bit size required to store makes it obvious that primitive data types cannot con-
the root of the polynomial is always equal or less verge after the polynomial’s 15th degree. Additionally,
than the highest bit size of the coefficients given to by employing big numbers, we are able to converge
that polynomial. up to a degree of 151, and the convergence time is
In the third experiment, all coefficient and con- recorded in milliseconds, demonstrating the speed of
stant values have been taken and processed with root convergence.
Applied Data Science and Smart Systems 305
Table 39.4 Polynomial root convergence with big numbers
S Degree Coefficients of the polynomial Constant Computed root Execution Evaluation Correct root
No. of the value of the using Newton time in of f(x) convergence
polynomials polynomial Raphson seconds
method
Abstract
The significance of the transistor and linked lists has not been widely recognized, despite its theoretical potential. In light of
the current state of collaborative setups, there is a pressing need among cryptographers to promptly pursue the simulation
of compilers. KamMone, a novel heuristic for massively multiplayer online role-playing games, presents a potential answer
to the aforementioned challenges. The performance investigation confirms three hypotheses, namely, the impact of Massively
Multiplayer Online on encrypted Application Programming Interfaces has diminished; the adjustability of heuristic through-
put has been observed; and the influence of Turing machines on system design has decreased. The authors intentionally elimi-
nate useful heuristics for application binary interface (ABI) and illustrate the importance of automating web browser ABI.
The process of hardware prototyping for trainable configurations is carried out using an overlay network within a meticu-
lously designed and thoroughly verified software environment. A series of novel experiments were conducted, wherein mul-
tiple facets were scrutinized, and the outcomes were thereafter investigated and evaluated in comparison to existing literature.
Keywords: Compact modalities, complexity theory, heuristics, wide-area networks, algorithm design, performance analysis
a
[Link]@[Link]
308 The influence of compact modalities on complexity theory
symmetric encryption. Additionally, a parallel incom- a similar manner, it is noteworthy to state that the
patibility is established among compilers. The authors utilization of extreme programming in the previously
direct their attention towards validating the capacity mentioned study (Culler and Kumar, 2003) differs
of agents and compilers to engage in interaction for from our methodology as the authors solely incor-
the purpose of attaining this objective. porate validated information into the framework
The ensuing sections of this work are structured in (Blum and Johnson, 1994; White and Hoare, 1992;
the following manner: Initially, the authors present a Sharma, 2004). Hence, despite substantial endeavors
justification for the indispensability of internet qual- in this domain, it is evident that the approach contin-
ity of service (QoS). To achieve this goal, the authors ues to be the favored framework among cyberneti-
provide data that challenges the belief that reinforce- cists (Thompson and Maruyama, 2001; Stearns and
ment learning and remote procedure calls (RPCs) are Gupta, 2005; Singh et al., 2019).
inherently incompatible. Expanding upon this line of A multitude of psychoacoustic and real-time sys-
argumentation, the authors contextualize their find- tems have been proposed in academic literature.
ings within the wider scope of extant scholarship Expanding upon this line of argumentation, the
in this specific domain. Furthermore, the authors authors put out an alternate methodology to tackle
have devised a comprehensive framework known as the aforementioned issue, which entails the regula-
KamMone to facilitate the implementation of large- tion of compiler enhancement (Jacobson, 1992).
scale technology, with the aim of attaining the afore- The magnitude of the significance of this revelation
mentioned objective. This framework challenges the for the complexity theory community remains to be
widely accepted notion that the cacheable method for ascertained. All of these proposed alternatives ques-
visualizing local-area networks adheres to a Zipf-like tion the fundamental assumption that standardized
distribution. The writers reach a conclusion in the protocols and IPv6 are naturally inherent in nature.
final analysis. Therefore, any comparisons to this specific piece of
work are erroneous. The field of software engineering
experiences expedited progress through the adoption
Related work
of novel technologies, resulting in cost reduction, time
The demand for the transistor was initially described savings, and improved quality. This study examines
by Edgar Codd (Jacobson, 1992; Kumar, 2001; the potential of technology developments to enhance
Thompson and Maruyama, 2001; Stearns and Gupta, the efficiency of software engineering processes, with
2005; Verma et al., 2019). In contrast, the intricacy a particular focus on mitigating phase-related chal-
of their methodology exhibits a quadratic increase lenges. The paper includes a section on the intersec-
in tandem with the expansion of the Internet. The tion of software engineering and artificial intelligence
study conducted by Bose et al. (Zhao et al., 1990) (AI), which is subsequently followed by sections on
proposes a potential use case for the establishment of emerging technologies in the field and an analysis of
consistent epistemologies. Nevertheless, it is impor- AI’s impact on software engineering. The paper con-
tant to acknowledge that the study lacks any explicit cludes with a summary of the findings (Uppal et al.,
mention of real implementation specifics (Codd and 2020, 2022). The authors proceed to conduct a com-
Wilson, 1996; Daubechies et al., 2003; Thakur et al., parison between the current technique and previous
2021). The experiment employs five discrete network methodologies in the field of lossless epistemology
setups with different arrangements of hosts, switches, (Blum et al., 1994; Culler and Kumar, 2003; Sato and
and data packets. The analysis of distributed denial- Wilson, 2004). Expanding on this line of argumenta-
of-service attacks incorporates various factors, tion, Nehru put forward a theoretical structure for
including detection time, round trip time, packet the practical application of “fuzzy” epistemologies.
loss, and attack type (Badotra and Panda, 2021). The However, he did not completely grasp the implications
utilization of data mining is prevalent in the process of exploring public-private key pairs at that particu-
of decision-making and the derivation of inferences lar time. The technique discussed above demonstrates
from information. This study investigates the tech- a greater level of vulnerability in comparison to our
niques, advantages, and disadvantages of several data own technique (Inder, 2020). In a similar manner, Adi
mining and machine language (ML) systems. This Shamir (Ritchie, 2005) presented a conceptual frame-
resource assists individuals in selecting the most suit- work for evaluating the simulation of the location-
able decision-making tools that align with their own identity split. Nevertheless, Shamir’s understanding of
requirements (Verma et al., 2019). Instead of design- the implications of augmenting agents, which would
ing “smart” archetypes, the authors address this issue facilitate the realization of voice-over-IP research
by utilizing knowledge-based technologies. This study at that time, was incomplete. While the authors do
builds upon a number of previous methodologies, not express any issues regarding Takahashi’s exist-
all of which have yielded unsatisfactory results. In ing methodology (Abiteboul, 1996), they argue that
Applied Data Science and Smart Systems 309
Hardware and software configuration interconnected Apples was more effective than their
Figure 40.2 illustrates the observed relationship simple distribution, in contrast to previous research
between distance and complexity, indicating that as findings. The authors subsequently recognize the
distance increases, complexity decreases. The discov- lack of success in past research endeavors aimed at
ery holds considerable ramifications for the domain facilitating this specific talent. Figure 40.4 presents a
of autonomous regulation. The inclusion of essen- visual representation of the median delay observed in
tial experimental information is often overlooked by the methodology under consideration, in contrast to
researchers; nevertheless, the authors of this work alternative methodologies.
have diligently incorporated such details in a com-
plete manner. The researchers conducted a hardware Dogfooding KamMone
prototype of a self-learning overlay network in order The authors have made a deliberate and focused
to demonstrate that trainable configurations do not attempt to offer an elaborate depiction of the setup
have the capability to impact the work of German for performance analysis. Therefore, the subsequent
physicist O. Johnson. The researchers incorporated emphasis will be placed on the analysis and interpre-
additional storage capacity in the concurrent cluster tation of the acquired results. The study encompassed
system in order to investigate technological aspects. four novel experiments done by the researchers. (1) In
Biologists have successfully achieved a 50% reduction the initial phase, the researchers proceeded with the
in the effective floppy disk size of MIT’s certifiable deployment of web browsers on a total of 74 nodes
testbed (Clark, 1991). To conduct an investigation on that were strategically scattered throughout a vast
the 10-node cluster, the researchers at the University network. Subsequently, they conducted a comprehen-
of California, Berkeley made the decision to remove a sive performance evaluation by comparing the out-
portion of the Ethernet connection, specifically 2kB/s, comes of this deployment with those obtained from
from the university’s network. In a similar vein, the locally executing public-private key pairs. (2) The
researchers extracted a 7 MB hard disk from the desk- energy efficiency of the ErOS, AT&T System V, and
top computers in order to examine the tape drive per- MacOS X operating systems was assessed. (3) The
formance of the XBox network (Shenker and Bose,
1991). German end-users successfully integrated
additional flash memory into UC Berkeley’s under-
water test-bed. The authors have made a deliberate
decision to omit specific findings in order to preserve
the confidentiality and anonymity of the participants.
Figure 40.3 presents a visual representation of the
anticipated complexity when comparing KamMone
with alternative methods.
The establishment of a suitable software environ-
ment necessitated a significant investment of effort;
nonetheless, the resultant solution demonstrated
substantial advantages. The technique was further
supported by the authors by the use of a stochas- Figure 40.3 Comparison of median latency between
tic runtime applet (Harris, 2004). The trials done the methodology and other methodologies
expeditiously demonstrated that the distribution of
Figure 40.2 Illustration of the phenomenon where dis- Figure 40.4 Comparison of expected complexity be-
tance increases as complexity decreases tween KamMone and other algorithms
Applied Data Science and Smart Systems 311
researchers conducted an investigation into the poten- to patched big multiplayer online role-playing games.
tial consequences that may arise from the utilization Furthermore, there exist disparities between the pres-
of opportunistically topologically partitioned flip-flop ent findings regarding median work factor obser-
gates as opposed to interruptions. (4) The research- vations and the results revealed in prior scholarly
ers conducted thorough testing of the application on investigations (Kaashoek et al., 2004), specifically in
their personal desktop computers, placing particular the influential study conducted by C. K. Kumar on
emphasis on monitoring the available hard disk space. massively multiplayer online role-playing games and
The studies were conducted in the absence of wide the observed throughput of NV-RAM.
area network (WAN) congestion or other discernible
performance constraints. Subsequently, we will com- Conclusion
mence an in-depth examination of the latter segment
of the experiments. The observed results cannot be This study aims to examine KamMone, a newly
solely attributed to mistakes made by the operator. developed atomic tool that is designed to optimize the
During the initial phase of installation, suitable ano- utilization of the location-identity split. One potential
nymization techniques were employed to ensure the constraint of KamMone is its present incapacity to
preservation of confidentiality for any sensitive data. offer e-commerce capability. The authors acknowl-
It is noteworthy to notice that red-black trees exhibit edge the presence of this constraint and express their
more consistent speed curves in the performance of intention to address it in their forthcoming research
USB keys when compared to microkernelized gigabit endeavors. The authors were motivated to investigate
switches. The authors have intentionally chosen to the potential suitability of reliable epistemologies. In
exclude these findings at the current time. a similar manner, the authors directed their atten-
The subsequent inquiry conducted by the author tion towards presenting substantiation to counter the
focuses on experiments (1) and (4), which were assertion that the predominant encryption algorithm
previously elucidated and visually represented in employed in the progression of electronic commerce
Figure 40.5. It is crucial to recognize that information functions with a temporal complexity of Ω(n). The
retrieval systems exhibit more consistent RAM space authors find no valid reason to exclude the utilization
curves when compared to micro kernelized systems. It of the application in easing the assessment of expert
is imperative to recognize that the process of software systems.
emulation involved the application of anonymization The utilization of the heuristic technique has the
techniques to safeguard the confidentiality of any capacity to efficiently tackle a wide range of chal-
sensitive data. The presence of software vulnerabili- lenges faced by modern cyberinformaticians. A sig-
ties within the system resulted in the manifestation nificant weakness of KamMone is to its inability to
of unforeseen phenomena witnessed over the course effectively visualize context free grammar, an area of
of the studies. In this study, the authors provide a concern that the authors intend to address in their
comprehensive analysis of experiments (1) and (4), forthcoming research endeavors. Furthermore, the
which were previously referenced. The cumulative authors have illustrated that the mobile algorithm,
distribution function depicted in Figure 40.4 exhibits which lacks widespread recognition, has a tempo-
a distinct heavy tail, suggesting a heightened degree of ral complexity of Θ(logn) when employed for XML
complexity. It is important to acknowledge that local- processing. On the other hand, it is commonly recog-
area networks demonstrate a lower level of discretiza- nized that the virtual algorithm is incapable of effec-
tion in the speed curves of floppy disks as compared tively replicating public private key pairs. While the
presented line of reasoning may seem unreasonable
at first glance, it fundamentally opposes the neces-
sity of providing mathematicians with evolutionary
programming. The authors express a desire to delve
deeper into the additional complexities related with
these issues in their forthcoming research endeavors.
References
Merriam, J. (2009). Where do constitutional modalities
come from - Complexity theory and the emergence of
intradoctrinalism. J. Juris, 3, 191.
Badotra, S. and Panda, S. N. (2021). SNORT based early
DDoS detection system using Opendaylight and open
Figure 40.5 Relationship between the expected block networking operating system in software defined net-
size of KamMone and latency working. Clus. Comput., 24, 501–513.
312 The influence of compact modalities on complexity theory
Blum, M. and Johnson, S. O. (1994). Visualization of fiber- Shastri, S. and Jones, S. (2004). SameVolt: Investigation
optic cables. J. Flex. Stochas. Models, 7, 41–55. of systems. Proc. Conf. Omnis. Linear-Time Inform.
Clark, D. (1991). The relationship between suffix trees and 67–74.
access points with Pud. Proc. IPTPS. 10–16. Shenker, S. and Bose, Y. (1991). The effect of linear-time
Codd, E. and Wilson, V. (1996). Deconstructing consistent models on hardware and architecture. Proc. Conf. Ho-
hashing. Proc PLDI. 47–42. mogen. Stable Archet.
Culler, D. and Kumar, O. (2003). On the deployment of Singh, J., Goyal, G., and Gupta, S. (2019). FADU-EV an
DHCP. Proc WWW Conf. 28–33. automated framework for pre-release emotive analy-
Daubechies, I., Stearns, R., and Thompson, D. (2003). Im- sis of theatrical trailers. Multimed. Tools Appl., 78,
provement of virtual machines. TOCS, 16, 159–199. 7207–7224.
Harris, H., Gayson, M., Jones, X., Bachman, C., and Cocke, Stearns, R. and Gupta, A. (2005). The impact of embed-
J. (2004). Enabling Lamport clocks and forward-error ded communication on operating systems. Proc. Conf.
correction. J. Class. Peer-to-Peer Relat. Algorith., 44, Com. Interact. Technol. 22–29.
20–24. Thompson, A. and Maruyama, Y. I. (2001). The influence
Inder, Shivani, Arun Aggarwal, Sahil Gupta, Sanjay Gupta, of constant-time technology on separated complexity
and Sanjay Rastogi. (2020). An integrated model of theory. Proc. OSDI.
financial literacy among B–school graduates using Thompson, N. (2003). On the understanding of write-
fuzzy AHP and factor analysis. The Journal of Wealth ahead logging. Proc. FPCA. 8–13.
Management. Uppal, M. and Gupta, D. (2020). The aspects of artificial in-
Jacobson, V. (1992). An exploration of thin clients with Fi- telligence in software engineering. J. Comput. Theoret.
libeg. Proc. Workshop Data Min. Knowl. Discov. 38, Nanosci., 17, 4635–4642.
55–62. Uppal, M., Gupta, D., and Mehta, V. (2022). A bibliomet-
Jones, G. (1993). Towards the evaluation of model check- ric analysis of fault prediction system using machine
ing. J. Knowl. Self-Learn. Theory, 64, 71–93. learning techniques. Challen. Opport. Deep Learn.
Kaashoek, M. F., Varadarajan, W. V., Dahl, O., Takahashi, Appl. Indus., 4, 109.
D., Lampson, B., Abiteboul, S., Hoare, C. A. R., and Ramamohan, Y., K. Vasantharao, C. Kalyana Chakravarti,
Cook, S. (2004). On the simulation of spreadsheets. and A. S. K. Ratnam. (2012). A study of data mining
Tech. Rep., 1173. tools in knowledge discovery process. International
Kumar, Z. (2001). A case for thin clients. Proc. SIGGRAPH. Journal of Soft Computing and Engineering (IJSCE),
996. 2(3): 191–194.
Ritchie, D. (2005) . A case for RAID. Proc. Symp. Stable White, Z. and Hoare, C. (1992). A case for checksums.
Archet. 42–47. Proc. Conf. Real-Time Inform. 44–52.
Sato, S. and Wilson, P. N. (2004) . Simulating Smalltalk us- Zhao, D., Takahashi, W., Anderson, G., Adleman, L., Gray,
ing peer-to-peer methodologies. Proc. MOBICOM. J., Li, F., and Martin, M. V. (1990). Deconstructing ac-
33–38. tive networks using addiblegad. Tech. Rep., 97.
Thakur, D., Singh, J., Dhiman, G., Shabaz, M., Gera, T. Zheng, W., Gupta, O., Subramanian, L., Thompson, P.,
(2021). Identifying major research areas and minor Smith, I., Sun, F., and Gupta, R. (2005). Synthesiz-
research themes of android malware analysis and de- ing DNS using omniscient technology. OSR, 94,
tection field using LSA. Complexity, 1–28. 52–66.
Sharma, L. (2004). Towards the emulation of I/O automata.
Proc. Symp. Self-Learn. Epistemol. 57–68.
41 Designing a hyperledger fabric-based workflow
management system: A prototype solution to enhance
organizational efficiency
Arjun Senthil K. S.a, Thiruvaazhi Uloli and Sanjay V. M.
Kumaraguru College of Technology, Coimbatotre, Tamilnadu, India
Abstract
Workflow management is crucial for organizations to operate efficiently and effectively. It helps businesses to streamline their
operations, reduce manual work, minimize errors, and improve overall productivity. The popular current solutions which
are paper based, or web application based requires technological upgrade. Blockchain not only fits the requirements in terms
of cost, scalability but also in addressing the security requirements owing to the inherent use of public key-based digital
signatures, hash and decentralized architecture being an integral part of its foundations. In this work we choose to build our
prototype solution based on hyperledger fabric which adds flexibility through the customizable consensus mechanism and
its permissioned nature makes the identity management and association of public key to identity a seamless task. We show
that this design better fits the requirements and enhances workflow management in the organizational context. By extension,
a similar design has the potential to efficiently meet several of organizational requirements where we need public key-based,
digitally signed, sustainable and scalable solutions built on top of distributed architecture.
a
arjunsenthil.19is@[Link]
314 Designing a hyperledger fabric-based workflow management system
the required approval. Once the mentor has given additional situations, it is also impractical. Physical
his or her approval, the student may speak with the signatures on paper are also susceptible to destruc-
department head. tion or loss while in transit. Delays and conflicts may
Before signing the withdrawal form/letter for the result from this.
semester examinations, along with any essential ref- A digitized signature, also known as a scanned
erences and messages, the department head will also signature, is the digital representation of a physical
verify the mentor’s approval, review the student’s situ- signature. It is simple to insert and provides a visual
ation, and give their consent after confirming that they representation of the signature in electronic docu-
have done so. The controller of examination must go ments. It may also be conveniently stored and accessed,
through the same procedure in order to approve the too. Digital signatures, however, offer a higher level
withdrawal request. The principal is then notified of of security than digitized signatures. Since, digitized
the request for final approval (Figure 41.1). signatures are simple to falsify or duplicate. Their
applicability in most situations may be constrained
Signing methods currently in use by the fact that they are not legally binding in many
Physical signatures on paper documents, digitized jurisdictions.
signatures (signature images that have been scanned), An extremely high level of security and non-repu-
and digital signatures are the three methods of signing diation is offered by a digital signature. Here, the
that are most frequently used in workflow manage- validity and integrity of the provided documents are
ment. These are typical in every industry. confirmed using cryptographic techniques. Digital
The most often used type of signature is a physi- signatures can be easily inserted into electronic docu-
cal one. It is challenging to copy or counterfeit. ments (Arya et al., 2021). They are also legally bonded
Additionally, it is simple to confirm by contrasting it in many jurisdictions. However, digital signatures
with the original document. However, it takes time require a digital certificate issued by a trusted third
because the signer needs to be there physically. In party called “certificate authority”. As not everyone
has the access to a digital certificate it can be a barrier
to adoption. In the event of a compromised digital
certificate or the loss of a private key, digital signa-
tures may become susceptible to attacks.
The current solutions available for workflow man-
agement are either expensive or difficult to learn. Thus,
making it necessary to develop a future-proof, easy to
use, low-cost and secure solution that addresses these
challenges.
their projects efficiently. It offers numerous tools to applications. Here is a quick summary of the three
increase efficiency, including task monitoring, time main blockchain technology generations.
management, team communication, and automation.
Document management software called Laserfiche First-generation blockchain: Bitcoin
enables companies to digitize their paper-based Blockchain technology is the foundation of Bitcoin,
records and streamline procedures. To increase decentralized digital money. It does away with the
effectiveness and productivity, it provides functions requirement for intermediaries like banks or govern-
including document capture, search, and retrieval in ments to handle transactions. The distributed ledger’s
addition to workflow automation. transparency and immutability are crucial in main-
taining the accuracy of all recorded transactions.
Conclusion Proof of work (PoW), a consensus technique, is used
These programs are made to assist companies with by Bitcoin to uphold network security and validate
workflow automation and streamlining the project new transactions. In order to add new blocks to the
management, and document digitization. However, blockchain, miners perform computing work to solve
these have a steep learning curve and are expensive to challenging mathematical riddles. It’s vital to remem-
scale in a big context. Additionally, there aren’t many ber that the scripting language used by Bitcoin has
choices for personalized modification. built-in restrictions and can only enable basic smart
contract features like multi-sign transactions. These
Blockchain technology agreements increase security because they involve
numerous parties and need for multiple signatures to
Blockchain in recent years be valid.
Due to its distinctive characteristics, blockchain tech-
nology has attracted a lot of attention recently. It is a Second-generation blockchain: Ethereum
decentralized, transparent, and immutable digital led- Developers can build and use smart contracts and
ger that can store information and transactions safely. decentralized applications on Ethereum, a unique
In other words, if new information is posted to the platform. It created a programming language that
blockchain, everyone can see the changes, making it can manage intricate contracts, making it the next
impossible for them to be changed. generation of blockchain technology. Initially,
The advantages of blockchain over traditional Ethereum validated transactions using a technique
databases and programs are numerous. A high level known as proof of work, but it eventually shifted
of integrity and availability is first and foremost guar- to a quicker and more effective technique known
anteed by the blockchain because it is very impossi- as proof of stake. The method is now quicker and
ble to hack or alter the data stored there. As a result, more environmentally friendly. Smart contracts on
transactions proceed more quickly and are less expen- Ethereum have made it possible to build a wide
sive because there are no longer any middlemen or range of cutting-edge applications, particularly in
intermediaries required. Finally, fraud is simpler to the area of decentralized finance. In conclusion,
spot and prevent since it offers a public and auditable Ethereum is a decentralized platform with cutting-
record of all transactions. edge capabilities that has opened the door for new
Blockchain is a decentralized, tamper-proof led- kinds of apps, especially in the area of decentralized
ger that can give all workflow participants access to finance.
transparency. Smart contracts, which can automati-
cally execute after certain criteria are satisfied, can be Third-generation blockchain: Hyperledger fabric
used with blockchain to automate certain phases in a This is a unique type of blockchain network designed
workflow. To save time, decrease manual errors, and specifically for companies and organizations. With
boost process effectiveness all at once. Blockchain more sophisticated features than other generations, it
can assist companies in decreasing processing, stor- is regarded as the most recent generation of block-
age, and data management expenses by eliminating chain technology. Its architecture for private networks
the need for intermediates and optimizing operations. which allows for restricted access, is a key feature. It
By offering a standardized platform for data inter- also has tight restrictions on who may do what and
change and workflow management, blockchain can offers a variety of options for reaching agreements on
make it easier for various systems and apps to operate transactions. Private channels are one special feature
together. that allows some users to conduct private transac-
tions. This is useful for keeping things private and
The different generations of blockchain secure. Another advantage is that it can be custom-
There have been multiple versions of blockchain ized for different uses, like managing supply chains
technology, each with unique characteristics and or handling trade finances. In summary, while Bitcoin
316 Designing a hyperledger fabric-based workflow management system
started blockchain and Ethereum introduced smart and cloud architecture. The article concludes by
contracts, hyperledger fabric is specifically made for reporting on a systematic literature review (SLR)
businesses, with powerful features that meet their examining the development of BMFA and impor-
needs. Each generation of blockchain technology adds tant conditions for implementation (Almadani et al.,
new features and expands what can be done with it. 2023).
A survey on distributed workflow management
Literature survey on existing blockchain solutions describes a workflow management system that uti-
An article exploring how secure electronic health lizes smart contracts on a blockchain to automate the
record (EHR) management provided by blockchain execution of tasks and the transfer of data between
technology has the potential to change healthcare. different parties in a distributed workflow. The sys-
Due to their centralized structures, traditional EHR tem is designed to be flexible, allowing users to define
systems have security flaws, but blockchain ensures workflows and modify them as needed, while also
tamper-proof records through decentralization and providing a high level of security through the use of
cutting-edge cryptography. Health care providers cryptographic protocols. Until now, companies that
can securely communicate information while giving facilitate workflows have been important in regu-
patients discretion over data access, which lowers lating the overall process by acting as choke points
operating costs and fraud. Decentralized, trustless or bottlenecks. However, it might be challenging to
transactions on the blockchain increase security impose the same regulatory obligations and respon-
and transparency. Data security is further improved sibilities on decentralized workflow facilitations
through identity and access management (IAM) sys- (Seppala et al., 2022).
tems and privacy-enhancing technologies (PET). In An article by Singh et al., investigates the advan-
order to provide safe, decentralized access manage- tages and obstacles associated with the implemen-
ment, the article introduces an IAM system that com- tation of blockchain technology in contemporary
bines blockchain, OAuth 2.0, and hyperledger fabric. business operations. The authors discuss the key
This system promises to enhance the privacy and characteristics of blockchain technology, including
integrity of patient data (Shrabani et al., 2024). decentralization, transparency, immutability, and
Fridgen et al. in his article proposed a solution security, and how they can be useful for processes
on cross-organizational workflow management such as supply chain management and financial
using blockchain. They emphasize that a tamper- transactions (Singh et al., 2020). They also provide
proof transaction history can represent a significant insights into different consensus mechanisms used
improvement for numerous workflows that span in blockchain technology, their advantages and dis-
organizational boundaries. This literature talks about advantages, and their suitability for different pro-
workflow on cross-organizational scale. It has taken cesses. They concluded by proposing a system based
a bank as its subject and works on how a blockchain on the consensus Practical Byzantine Fault Tolerance
powered cross-organizational workflow tool will help (PBFT) explaining its versatility (Viriyasitavat et al.,
improve efficiency. To develop this workflow environ- 2021).
ment, it follows the design science research (DSR) Evermann’s article recommends that a semi appli-
approach. DSR tries to solve organizational problems cation that resides both on- and off-chain might be a
that are already identified through a build and evalu- great way to mitigate the flaws of the PoW system.
ate process (Fridgen et al. 2018). They too suggest the PBFT consensus acknowledg-
M. S. Almadani et al. in his article explains the ing that it’s hard to scale but gives finality to transac-
importance of multi-factor authentication (MFA) in tions and has low latency. The proposal is to integrate
enhancing security is discussed in the article, particu- Byzantine Fault Tolerance (BFT)-based blockchains
larly in distributed systems like blockchain networks. into workflow management systems in order to iden-
It describes the three techniques of MFA—knowl- tify potential design problems and assess their impact
edge, possession, and inheritance—as well as how it on both the systems themselves and their users. They
uses distinctive authentication components. Due to its also emphasize the importance of clean user interface
dependability and immutability, blockchain technol- and user education for wide acceptance of the service
ogy is suggested as a secure alternative to centralized (Evermann et al., 2021).
authentication in distributed systems, highlighting A survey on workflow management on BFT
its weakness. In his work it discusses authentication explains the multiple blockchains built around BFT
methods based on blockchains and how they moved algorithms are proven to be more efficient than the
away from centralized credential storage and toward other available solutions. It also provides immediate
decentralized ledger storage. Additionally, it empha- consensus. However, it does not scale well to large
sizes the need to improve MFA for blockchain net- networks since the number of nodes will increase and
works by mentioning the integration of blockchain the execution time will also increase in return. The
Applied Data Science and Smart Systems 317
increased requirement for processing power is the security by optimizing digital signature algorithms.
major drawback of the PoW system. Moreover, it also Integrating digital signatures with identity authenti-
has increased latency and there is no finality of con- cation or timestamps can provide multi-dimensional
sensus. Whereas BFT-SMART utilizes a PBFT-based security and safeguard information non-repudiation
ordering mechanism that eliminates the latency, lack in the blockchain from a broader perspective (Fang
of finality, and computational demands of the PoW et al., 2020).
consensus. But it requires fully connected nodes and A survey on the preservation of digital signatures.
perfect communication overhead (Evermann et al., Traditional infrastructures announce the authenticity
2019). of key pairs and digital signatures using digital certifi-
Another article by Evermann et al., explores the cates, which are given by certification authority like
advantages and obstacles of using blockchain technol- Adobe. Digital signatures, blockchain, keys, encryp-
ogy for workflow management. The authors suggest tion, authenticity, and trust are all terms that are used
that blockchain technology can improve workflow in this paper to argue that the hash functions of the
management systems by providing a decentralized, blockchain provide a superior technique for main-
transparent, and secure solution. However, there are taining signatures than digital certificates (Thompson,
challenges such as scalability, privacy, and regulatory 2020).
compliance that need to be considered. The paper A decentralized web application for digital docu-
provides insights into different types of blockchain ment verification using Ethereum blockchain-based
technology, the importance of smart contracts, and technology in P2P cloud storage. The goal of this
recommendations for successful implementation application is to enhance the verification process by
(Evermann et al., 2019). making it more transparent, accessible, and audit-
A solution that included proof of storage and proof able. The proposed model utilizes various tech-
of existence. By including all of these features, CDAC niques such as public/private key cryptography,
created ProveDoc, a solution that proves the tempo- online storage security, digital signatures, hashing,
ral existence of any digital document, authenticates peer-to-peer networks, and proof of work, making it
its content, confirms the document’s origin, ensures faster and more convenient for any organization or
that its timestamp and hash cannot be altered ret- authority to verify uploaded documents with a sin-
roactively, and gets around the problem of storing gle click. Each document is also assigned an appro-
large data directly in blockchain with the aid of PoS priate hash value. By addressing the limitations of
(Chiliveri et al., 2019). traditional document verification methods, our pro-
A solution delays in payments and human error posed model effectively meets all the requirements
in cash flow management for construction projects for a digital document verification system (Imam et
continue, necessitating the use of digital technologies. al., 2021).
As a decentralized solution, blockchain automates A blockchain-based solution for storing and shar-
processes and improves transparency. Current solu- ing records across institutions, ensuring security and
tions have drawbacks like centralization and labori- integrity using a consortium blockchain. By combin-
ous data entry, such as cash flow-based 5D BIM and ing a storage server with a blockchain, secure docu-
web-based management systems. To overcome these ment storage is created. Smart contracts are utilized
difficulties, a networked financial management sys- to enable cross-institutional sharing of educational
tem employing chaincode and the hyperledger fabric records, with the consortium blockchain’s smart con-
is being developed. It offers a “proof of concept” solu- tracts regulating document exchange permissions
tion for all project stakeholders, classifying roles for and processes between institutions. Additionally, an
various parties and making it possible to trace finan- anti-tampering inspection method is employed to pro-
cial transactions over the course of a project. This tect the records stored in the storage server (Li et al.,
adaptable system takes into account different pro- 2019).
curement strategies, boosting trust and transparency
in the financial administration of building projects Design of solution
(Elghaish et al., 2022).
A digital signature scheme for non-repudiation. Comparison of available options
They analyzed various digital signature schemes Paper-based workflow management methods use
used in blockchain systems over the past few years actual paper documents and manual procedures. Due
and found that digital signature technology can fulfill to the possibility of lost, damaged, or missing papers,
specific application requirements of blockchains and this process can be time-consuming and error-prone.
meet security needs in diverse situations. The findings On the other hand, web applications are computer
of this study can aid in the design of digital signa- programs that may be accessed through a web browser.
ture schemes for blockchain and enhance blockchain They make it simpler to manage jobs and monitor
318 Designing a hyperledger fabric-based workflow management system
Summary
While Ethereum is a public blockchain platform
appropriate for decentralized apps and coin creation
utilizing smart contracts. Hyperledger fabric is made
for usage in business applications that need authentic-
ity, confidentiality, integrity, and scalability. Businesses
who want a private blockchain network for secure
transactions and data sharing can use it because of
its permissioned approach. For identity management,
hyperledger fabric is a fantastic alternative and is
used in sectors including finance, healthcare, and sup-
ply chain management. As a result, we developed a
blockchain-based workflow management application
using hyperledger fabric. The purpose of this program
is to facilitate a simple and secure workflow among Figure 41.2 Prototype application basic design
network users.
“[Link](args[0], dataInBytes)” increased security, and cost savings. This study adds
to the conversation about workflow management’s
where dataInBytes is a struct containing the signature ongoing pursuit of innovation and quality.
and timestamp and args[0] is a unique key which is
later used to retrieve the signature.
References
Developing the application Reijers, H. A., Vanderfeesten, I., and Van Der Aalst, (2016).
The frontend of the application is made simple, The effectiveness of workflow management sys-
intuitive and easy to use with the various libraries of tems: A longitudinal study. Int. J. Inform. Manag.,
[Link]. Once a user sends a document for approval, 36(1), 126–141. [Link]
the document is stored in a storage and the informa- FOMGT.2015.08.003.
tion regarding it is stored in a MongodB database. Wu, D. T. Y., Barrick, L., Ozkaynak, M., Blondon, K., and
With the network and chain code in place the server Zheng, K. (2022). Principles for designing and devel-
can send the necessary data to the network after the oping a workflow monitoring tool to enable and en-
document gets approved by the mentor. The network hance clinical workflow automation. Appl. Clin. In-
format., 13(1), 132–138.
can now store the necessary data to be later retrieved
Shrabani, S., Karforma, S., Bose, R., Roy, S., Djebali, S., and
for verification. Bhattacharyya, D. (2024). Enhancing identity and ac-
cess management using hyperledger fabric and OAuth
Conclusion 2.0: A blockchain-based approach for security and
scalability for healthcare industry. Internet of Things
In conclusion, this study emphasizes how crucial Cyber-Phy. Sys., 4.
workflow management is to maintain an organiza- Fridgen, G, Urbach, N., Radszuwill, S., and Utz, L. (2018).
tion’s efficacy and efficiency. It is clear that workflow Cross-organizational workflow management using
management is essential for automating processes, blockchain technology – towards applicability, au-
reducing manual labor, reducing human error, and ditability, and automation. Proc. Ann. Hawaii Int.
increasing productivity in general. Conf. Sys. Sci., 2018, 3507–3516. doi: 10.24251/
The existing solutions, which are frequently paper- HICSS.2018.444.
based or dependent on online applications, must be Almadani, Mwaheb S., Suhair Alotaibi, Hada Alsobhi, Omar
K. Hussain, and Farookh Khadeer Hussain. (2023).
upgraded technologically in order to fully meet cur-
Blockchain-based multi-factor authentication: A sys-
rent expectations. Given that blockchain technology tematic literature review. Internet of Things: 100844,
satisfies important criteria including cost effective- 23, [Link]
ness, scalability, and security, it is seen as a promis- Hukkinen, Taneli, Juri Mattila, and Timo Seppälä. (2017).
ing solution. Its inherent usage of public key-based Distributed workflow management with smart con-
digital signatures, cryptographic hash functions, and tracts. No. 78. ETLA Report.
a decentralized architectural foundation promote this Belhi, Abdelhak, Houssem Gasmi, Abdelaziz Bouras, Be-
alignment. laid Aouni, and Ibrahim Khalil. (2021). Integration
We have developed a hyperledger fabric-based pro- of business applications with the blockchain: Odoo
totype solution as part of this investigation. By allow- and hyperledger fabric open source proof of concept.
ing for configurable consensus processes, this decision IFAC-PapersOnLine 54(1): 817–824.
Evermann, Joerg, and Henry Kim. (2021). Workflow man-
gives our workflow management system more flex-
agement on proof-of-work blockchains: Implica-
ibility and makes it easier to link public keys to iden- tions and recommendations. SN Computer Science
tities inside its permissioned framework. 2: 1–22.
The implications go beyond this particular context, Evermann, Joerg, and Henry Kim. (2019). Workflow
indicating that comparable blockchain-based systems Management on BFT Blockchains. arXiv preprint
could be used to satisfy a range of organizational arXiv:1905.12652, [Link]
needs. This is especially true in situations when the iv.1905.12652.
need for public key-based, digitally verified, resilient, Evermann, Joerg, and Henry Kim. (2020). Workflow man-
and scalable solutions collide with distributed archi- agement on BFT blockchains. Enterprise Modelling
tecture principles. and Information Systems Architectures (EMISAJ) 15:
The incorporation of blockchain technology into 14–1.
Chiliveri, S., Grandhi, J., Uttam Patil, M., Lakshmi Eswari,
workflow management presents a practical route
P. R., and Ethirajan, M. (2019). ProveDoc: A block-
to operational efficiency given all the technological chain based proof of existence with proof of storage.
developments reshaping the organizational land- Proc. 2019 Int. Conf. Inform. Technol. ICIT 2019.,
scape. Organizations have compelling motivations 239–244, doi: 10.1109/ICIT48102.2019.00049.
to investigate and deploy blockchain-based solu- Arya, Resham, Jaiteg Singh, and Ashok Kumar. (2021). A
tions due to the promise of improved operations, survey of multidisciplinary domains contributing to
Applied Data Science and Smart Systems 321
affective computing. Computer Science Review 40: omnichannel business. Enterp. Inform. Sys., 14(2),
100399. 243–265. [Link]
Elghaish, Faris, Farzad Pour Rahimian, M. Reza Hos- 40392.
seini, David Edwards, and Mark Shelbourn. (2022). Thompson, Stephen. (2017). The preservation of digital sig-
Financial management of construction projects: Hy- natures on the blockchain. See Also 3, DOI: https://
perledger fabric and chaincode solutions. Automation [Link]/10.14288/sa.v0i3.188841.
in Construction 137: 104185. Imam, I. T., Arafat, Y., Alam, K. S., and Aki, S. (2021). DOC-
Fang, W., Chen, W., Zhang, W., Pei, J., Gao, W., and BLOCK: A blockchain based authentication system for
Wang, G. (2020). Digital signature scheme for in- digital documents. Proc. 3rd Int. Conf. Intel. Comm.
formation non-repudiation in blockchain: a state Technol. Virt. Mob. Netw. ICICV 2021, 1262–1267.
of the art review. EURASIP J. Wirel. Comm. Netw., doi: 10.1109/ICICV50876.2021.9388428.
2020(1), 1–15. doi: 10.1186/S13638-020-01665-W/ Li, H. and Han, D. (2019). EduRSS: A blockchain-based ed-
TABLES/2. ucational records secure storage and sharing scheme.
Singh, J., Goyal, G., and Gill, R. (2020). Use of neuro- IEEE Acc., 7, 179273–179289. doi: 10.1109/AC-
metrics to choose optimal advertisement method for CESS.2019.2956157.
42 Exploring Image Segmentation Approaches for Medical
Image Analysis
Rupali Pathak1,a, Hemant Makwana2 and Neha Sharma1
1
Prestige Institute of Engineering Management and Research, Indore, India
2
Institute of Engineering and Technology, DAVV, Indore, India
Abstract
The research provides a review of segmentation methods for medical imaging. The article provides a comparative study of
edge-based, region-based and energy-based methods of image segmentation. Many medical images have different levels of
intensity because of flaws in the object or technical limitations. Some segmentation techniques have limitations, like being
stuck in local minima or producing over-segmented images. Medical image segmentation is still difficult because of noise,
poor contrast, and huge variations in intensity level. Interactive medical image segmentation is also needed for better results
that can be solved by taking user input into account while segmenting an image. Active contour techniques assume that
global intensity may be used to define an image. Some approaches deal with intensity inhomogeneity in the same way that
the region-scalable fitting (RSF) model does. The method was compared to Chan-Vese (CV), RSF, and local and global in-
tensity fitting (LGIF). The hybrid region-based active contour (HRBAC) approach may be useful in addressing the intensity
of inhomogeneity. It also accelerates segmentation as compared to local region-based active contour (LRBAC). However,
HRBAC has limitation of contour initialization sensitivity and parameter sensitivity. Improved HRBAC model is also ex-
plored in this paper to deal with challenges of medical image segmentation mentioned above. The new function used in this
model will leverage global and local data to quickly get the correct response. The method combines energy functional-driven
curve generation with a level set framework for medical picture segmentation. Lattice Boltzmann approach is used to make
segmentation process fast.
Keywords: Active contour model, edge-based method, energy-based method, intensity heterogeneity, medical image segmen-
tation, region-based method
rpathak@[Link]
a
Applied Data Science and Smart Systems 323
2021). There are several solutions to the CV model’s these models (Wang et al., 2014). Both models’ local
flaws. For two models that are regionally comparable intensity fitting terms were often combined to see how
(Vese, 2002). Any region-based segmentation energy, they affected curve creation in various places. In this
according to Lankton and Tannenbaum (Hemalatha instance, global intensity information takes prece-
et al., 2018), can be recast locally. Active contour dence. The contour is drawn to the boundaries of the
energy is being used to segment objects with variable item and then stopped. The use of local image con-
statistics. However, they are CPU-intensive, which trast modifies its weight (Wang et al., 2010; Memon
suggests an initial contour at the object’s edges. Zhang et al., 2020). At weak object borders, the global
et al. (2010) and Li et al. (2020) provide region-based intensity force leads the contour to diverge. More dis-
active contour models for dealing with intensity inho- criminative energy functions are necessary to increase
mogeneity. Local binary fitting (LBF) and region- model performance. The development of contours is
scalable fitting (RSF) are the most frequently used governed by discriminant and fitting. New region-
models by Li et al. The LBF model makes use of local scalable discriminants and energy functionalities were
image data. The RSF model makes use of local inten- introduced. While its counterpart phrase describes
sity data. Both models may be used at the same time intensity, this word differentiates between the back-
(Chuanjiang et al., 2012) (Figure 42.1). ground and the foreground. The energy is computed
using level set regularization. It manages intensity
Related work inhomogeneity better than conventional regional and
regional-scalable models due to its more flexible ini-
Many real-world images exhibit intensity heteroge- tialization. The new model trumps the old. To begin
neity. It’s widespread in medical images like X-rays with, the new energy functional provides proper fore-
and MRIs (MR) (Vovk et al., 2007). Radiofrequency ground/background separation. The method keeps
coils or acquisition processes generate inhomogene- both local and global data. There is now a sensitive
ity un-MR images. The intensity of the same tissue contour. These photographs highlight the accuracy
changes with time. In CT and ultrasound pictures, and longevity of the process. Piovano et al., employed
non-uniform beam attenuation generates similar convolutions to accelerate piecewise smooth seg-
issues. A new RSF model (Brox et al., 2009) inves- mentation. It deals with picture intensity and spatial
tigated the effect of intensity heterogeneity on seg- variations directly. Instead of calculating a piecewise
mentation. The RSF model makes use of locally smooth model as suggested in Chunming et al. (2009),
fluctuating data that fluctuate geographically (Zhi et there is less reliance on initial curve location. On the
al., 2015; Koshki et al., 2021). This is because the RSF other hand, the Geodesic Active Contour and Chan-
model may effectively segment using local area infor- Vese models are combined in this model. The geodesic
mation, namely the local intensity mean. Some recent intensity fitting (GIF) model was created. Later, two
suggestions deal with intensity inhomogeneity in the models emerged: the GGIF global model and the LGIF
same way that the RSF model does. However, some local model. The GGIF model is intended for pictures
approaches need setup, which limits their applicabil- that are uniform in size. The LGIF model accounts for
ity. Over-reliance on the contour’s start location is a intensity inhomogeneity. The new function will lever-
major flaw in local information models, solving Euler- age global and local data to quickly get the correct
Lagrange equations reduces energy functions. Local response. The CV model provides global data. The
and global intensity fitting energies are included in local information is explained by the energy para-
digm in Wan et al. (2018) using inter-fitting weights to
avoid computationally costly and erroneous segmen-
tation. Many image processing and computer vision
applications make use of it. Active Contour Model
and fuzzy C-means (FCM) are two well-known image
segmentation methods. Medical image segmentation
is still difficult because of noise, poor contrast, and
a lot of variation in intensity. A hybrid region-based
contour model (HRBAC) is also deals with intensity
inhomogeneity with efficiency (Liu et al., 2014; Xu
et al., 2014). This approach combines the advantages
of global and local region based active contour mod-
els. Localizing region based active contour (LRBAC)
and GIF are also handle intensity inhomogeneity but
may be computationally expensive or sensitive to
Figure 42.1 The segmentation outcomes initialization.
324 Exploring Image Segmentation Approaches for Medical Image Analysis
Findings
Experiments have proven that automatic segmenta-
tion methods do not give correct analysis as needed
for medical images.
Figure 42.3 Real-world medical images are used to test the procedure. Column 1 contains the original images and con-
tours. Column 2 has the final outlines. Column 3 contains photos that have been adjusted for bias. Column 4 contains
the estimated bias fields
Applied Data Science and Smart Systems 325
Figure 42.4 Alternative approaches are compared. Column 1 displays the original pictures and initial outlines. The C-V
model is shown in column 2. There are columns 3: Li’s method. The LGDF model may be found in column 4. Column
5 – In this scenario, the method should be followed
moving as you move from one area to another is associated with segmentation of medical images,
eliminated. Using statistical data, the model creates energy-based models perform better than alterna-
a region-halting function. The addition of a statisti- tives. New model based on regions for medical image
cal region allows for the expansion of this function. segmentation Gaussian distributions with varying
If any image has fuzzier edges, this model will out- means and variances are used to establish statistics
perform the edge-based technique. When it comes to of picture intensities for objects in local areas. As a
image segmentation, region-based models are pre- result, it can improve segmentation accuracy. Both
ferred over edge-based models. This is since apply- synthetic and real-world medical imagery work effec-
ing region-based models has no constraints, whereas tively. The model can be multiphase, allowing it to
applying edge-based models does. When examined understand more complex medical images with dif-
side by side, region-based models usually outperform ferent levels of intensity.
edge-based models. Because they assume that all ele-
ments of a picture are the same, the standard region- References
based models suggested for binary images may not
perform as well for images with intensity inhomo- Qu, Xiaoxia, Jian Yang, Danni Ai, Hong Song, Luosha
geneity. These models assume that there are no dis- Zhang, Yongtian Wang, Tingzhu Bai, and Wilfried
Philips. (2017). Local directional probability optimi-
cernible differences across image regions. As a result
zation for quantification of blurred gray/white mat-
of the preceding contour, the developing curve may ter junction in magnetic resonance image. Frontiers in
become caught in local minima. Because computing Computational Neuroscience 11: 83.
the standard intensities both inside and outside the Kumar, Naresh. (2010). Gradient Based Techniques for the
contour takes time, the CV technique is inappropriate Avoidance of Oversegmentation. Proceedings of the
for application in circumstances requiring fast pro- BEATS, 1–6.
cessing. The longer it takes to compute the results, the Ge, Qi, Liang Xiao, Jun Zhang, and Zhi Hui Wei. (2012).
less suitable the method is for the speedy processing An improved region-based model with local statistical
required. Standard region-based models do not per- features for image segmentation. Pattern Recognition
form as well on binary images as they do on images 45(4): 1578–1590.
with relative intensity variation. By drawing on data Guo, M., Zhaobin, W., Yide, M., Weiying, X. (2013). Re-
view of parametric active contour models in image
from surrounding images, the LBF model improves
processing. J. Converg. Inform. Technol., 8, 248–258.
previously proposed strategies for segmenting images 10.4156/jcit.vol8.issue11.28.
with high intensity inhomogeneity. This allows the Sun, L., Meng, X., Xu, J., and Tian, Y. (2018). An image
LBF model to merge data from multiple pictures into a segmentation method using an active contour model
single image. This outcome is more plausible because based on improved SPF and LIF. Appl. Sci., 8(12),
the model uses locally derived visual information. The 2576.
model’s ability to separate images using information Liu, S., and Peng, Y. (2012). A local region-based Chan–
from similar photos enables this. The main reason for Vese model for image segmentation. Patt. Recogn.,
its inclusion is the desire to include the Gaussian ker- 45(7), 2769–2779. ISSN 0031-3203.
nel function, even though it is quite good at segment- Krupinski, E. A. (2010). Current perspectives in medical
ing images with inhomogeneous intensities. image perception. Atten. Percept. Psychophys., 72(5),
1205–1217. doi: 10.3758/APP.72.5.1205. PMID:
20601701; PMCID: PMC3881280.
Conclusion Hemalatha, R., T. Thamizhvani, A. Josephin Arockia
Dhivya, Josline Elsa Joseph, Bincy Babu, and R. Chan-
The HRBAC approach can be useful in addressing drasekaran. (2018). Active contour based segmenta-
the intensity of inhomogeneity. It also accelerates seg- tion techniques for medical image analysis. Medical
mentation as compared to LRBAC. HRBAC outper- and Biological Image Analysis 4(17): 2.
forms the CV model and LRBAC in terms of intensity Li, Y., Cao, G., Wang, T., Cui, Q., and Wang, B. (2020). A
inhomogeneity and noise resilience. The energy func- novel local region-based active contour model for im-
tional in the model is non-convex, having local min- age segmentation using Bayes theorem. Inform. Sci.,
ima, and hence sensitive to contour initialization. In 506, 443–456. ISSN 0020-0255.
improved HRBAC use of lattice boltzamnn method Chuanjiang, H., Wang, Y., and Chen, Q. (2012). Active con-
make it fast than others models. In this method tours driven by weighted region-scalable fitting en-
results are same irrespective of the initial contour ergy based on local entropy. Sig. Proc., 92, 587–600.
10.1016/[Link].2011.09.004.
position that’s make it interactive. From this study,
Chunming, L., Li, F., Kao, C.-Y., and Xu, C. (2009). Image seg-
we can conclude that image segmentation methods mentation with simultaneous illumination and reflec-
based on region-based are preferable over edge-based tance estimation: An energy minimization approach.
models. For medical images with varying intensities, ICCV., 702–708. 10.1109/ICCV.2009.5459239.
these models perform better. To deal with challenges
Applied Data Science and Smart Systems 327
Xu, H., Liu, T., and Wang, G. (2014). Hybrid geodesic re- Chan, and Vese, (2001). Active contours without edges.
gion-based active contours for image segmentation. IEEE Trans. Image Proc., 10(2), 266–277.
Comp. Elec. Engg., 40(3), 858–869. Wang, L., He, L., Mishra, A., and Li, C. (2009). Active con-
Wan, M., Gu, G., Sun, J., Qian, W., Ren, K., Chen, Q., and tours driven by local Gaussian distribution fitting en-
Maldague, X. (2018). A level set method for infra- ergy. Sig. Proc., 89(12), 2435–2447.
red image segmentation using global and local in- Li, C., Huang, R., Ding, Z., Gatenby, J., Metaxas, , and
formation. Remote Sens., 10(7), 1039. [Link] Gore, (2011). A level set method for image segmenta-
org/10.3390/rs10071039. tion in the presence of intensity inhomogeneities with
Koshki, A. S., Ahmadzadeh, M. R., Zekri, M., Sadri, S., and application to MRI. IEEE Trans. Image Proc., 20(7),
Mah-moudzadeh, E. (2021). A level-set method for in- 2007–2016.
homogeneous image segmentation with application to Kass, M., Witkin, A., and Terzopoulos, D. (1988). Snakes:
breast thermography images. IET Image Proc., 15(7), active contour models. Int. J. Comp. Vis., 1(4), 321–
1439–1458. 331.
Modi, Nandini, and Jaiteg Singh. (2021). A review of An, J., Rousson, M., and Xu, C. (2007). Γ-convergence ap-
various state of art eye gaze estimation techniques. proximation to piecewise smooth medical image seg-
Advances in Computational Intelligence and Com- mentation. Med. Image Comp. Comp-Ass. Interven.-
munication Technology: Proceedings of CICT 2019: MICCAI, 4792, 495–502.
501–510. Vese, and Chan, (2002). A multiphase level set framework
Zhi, X., Ting-Zhu, H., Hui, W., and Chuanlong, W. (2015). for image segmentation using the Mumford and Shah
Variant of the region-scalable fitting energy for image model. Int. J. Comp. Vis., 50(3), 271–293.
segmentation. J. Opt. Soc. Am. A, 32, 463–470. Zhang, K., Song, H., and Zhang, L. (2010). Active contours
Memon, A. A., Soomro, S., Tanseef Shahid, M., Munir, driven by local image fitting energy. Patt. Recogn.,
A., Niaz, A., Choi, K. M. (2020). Segmentation of 43(4), 1199–1206.
intensity-corrupted medical images using adap- Liu, Tingting, Haiyong Xu, Wei Jin, Zhen Liu, Yiming Zhao,
tive weight-based hybrid active contours. Comput. and Wenzhe Tian. (2014). Medical image segmenta-
Mathemat. Methods Med., 2020, 14. [Link] tion based on a hybrid region-based active contour
org/10.1155/2020/6317415. model. Computational and mathematical methods in
Wang, H., Ting-Zhu, H., Xu, Z., and Wang, Y. (2014). An medicine 2014.
active contour model and its algorithms with local Vovk, U., Pernuš, F., and Likar, B. (2007). A review of meth-
and global Gaussian distribution fitting energies. In- ods for correction of intensity inhomogeneity in MRI.
form. Sci., 263, 43–59. ISSN 0020-0255. IEEE Trans. Med. Imag., 26(3), 405–421.
Singh, Jaiteg, and Nandini Modi. (2019). Use of informa- Brox, T. and Cremers, D. (2009). On local region mod-
tion modelling techniques to understand research els and a statistical interpretation of the piecewise
trends in eye gaze estimation methods: An automated smooth Mumford-Shah functional. Int. J. Comp. Vis.,
review. Heliyon. 5(12). 84(2), 184–193.
Wong, and Chung, (2005). Bayesian image segmentation Wang, X., Huang, ,and Xu, H. (2010). An efficient lo-
using local iso-intensity structural orientation. IEEE cal Chan-Vese model for image segmentation. Patt.
Trans. Image Proc., 14(10), 1512–1523. Recogn., 43(3), 603–618.
Li, C., Kao, , Gore, ,and Ding, Z. (2008). Minimization of
region-scalable fitting energy for image segmentation.
IEEE Trans. Image Proc., 17(10), 1940–1949.
43 Design and performance analysis of electric shock
absorbers
Jenish R. P.a and Surbhi Gupta
Department of Computer Science and Engineering, Chandigarh University, Punjab, India
Abstract
Modern vehicles have attained remarkable dynamics, largely attributed to the reduced vibration emanating from the pow-
ertrain and the damping effects of shock absorbers. However, the effectiveness of conventional shock absorbers is curtailed
by their inherent mechanical structure, which generally confines them to fixed output levels. Notably, luxury cars distinguish
themselves by integrating shock absorbers equipped with variable outputs to optimize comfort and performance. This paper
delves into a pioneering realm: electric shock absorbers with variable outputs, a concept with universal applications across all
vehicle types. The core objective is to unravel the potential of this novel suspension technology and its transformative impact
on vehicle dynamics. By exploring the territory of electric shock absorbers with variable outputs, this research contributes
to an evolving field that is set to redefine vehicular comfort and handling. The proposition of applying this concept to all
vehicles opens up avenues for a more standardized and enhanced driving experience, transcending the confines of luxury car
segments. Through a comprehensive exploration of electric shock absorbers with variable outputs, this study embarks on a
journey to revolutionize the realm of vehicle dynamics, promising an era of superior ride quality and enhanced maneuver-
ability.
Keywords: Shock absorber, hydraulic damper, hydraulic valve, bode plot, PID controller
jenishraj97@[Link]
a
Applied Data Science and Smart Systems 329
Foundation
Matlab is one of the dominant analysis software
where we can perform our concept analysis theo-
retically with high precision, here we will be using
Figure 43.3 Valve after rotation Simulink as a floor to analyses our model which is an
tool in the Matlab.
(1)
(2)
(3)
(4)
role in the stability of the vehicle and the exact ride Output equation
frequency will determines the comfort (Tran et al.,
2022) (Figure 43.4). (6)
Ride frequency of the front and the back wheel can
be controlled by changing the damping ratio which is Using these equations, MATLAB model is made and
done using the valve rotation. When the delay between separated in to various subsystem for further analysis
the front and the back wheel is reduced damping will (Figure 43.6).
happen in the same proportion which can give a good
vehicle dynamic (Figure 43.5).
Result
Also, various dynamics of the vehicle is considered
for designing the PID controller so that the perfor- Analysis of this model is made with some of the base
mance in dynamic condition can be made smooth and assumption which are, the mass is set to be 200 kg
effective. and spring rate is set to be 32 N/mm, damping ratio
Applied Data Science and Smart Systems 331
was changed in four scenarios which is because when discussed when the valve rotates the damping will be
the valve rotates the hole size will vary leading to changing. The results prove that we can change the
change in damping ratio damping based on load applied and the required com-
From Figure 43.7 we can see that damping fre- fort of the passenger. In practical condition damping
quency of the various damping condition as we will not be constant throughout the motion based on
the load factor it will vary for matching the required suspension system. This system can detect and react
damping, we can rotate the valve and try to match to various driving conditions, ensuring optimal per-
the damping value which is required. Dampers are formance while maintaining consistent ride quality.
normally inside the spring which has compression Energy efficiency: The integration of low-voltage
and rebound, however, during the rebound its good stepper motor technology minimizes power con-
if the spring is capable of returning faster. If there is sumption. This energy-efficient design aligns with the
over damping while returning it may cause failure of industry’s push toward sustainable solutions without
the shock absorbers, here we can change the damp- compromising on performance.
ing even while rebound which will helps the spring to Universal application: While luxury cars have his-
return faster. torically featured variable output shock absorbers,
this technology’s universal application opens doors
Features for standardizing advanced suspension systems across
The concept of electric shock absorbers with variable a wider range of vehicles. This democratization of
outputs introduces a host of features that promise to enhanced dynamics marks a significant shift in the
revolutionize the realm of vehicle dynamics and rede- automotive landscape.
fine the driving experience across all vehicle types.
These features not only enhance comfort but also Conclusion
contribute to unprecedented levels of stability, con-
trol, and adaptability: In the pursuit of refining vehicle dynamics and rei-
Dynamic damping control: Electric shock absorb- magining the driving experience, the concept of elec-
ers with variable outputs enable real-time adjust- tric shock absorbers with variable outputs emerges
ments to the damping characteristics. This dynamic as a pivotal innovation. This research delves into
control allows the vehicle’s suspension system to uncharted territory, unearthing the potential to trans-
adapt instantly to changing road conditions, provid- form how vehicles interact with the road and how
ing a smoother ride and improved handling. passengers perceive the journey.
Tailored ride comfort: The ability to adjust damp- The features outlined above underscore the trans-
ing ratios based on load conditions, road surfaces, formative impact of this technology, transcending
and driving speeds offers personalized ride comfort to traditional limitations and offering a comprehensive
passengers. This feature transcends the one-size-fits- solution to the challenges posed by varying road con-
all approach of traditional shock absorbers, ensuring ditions and driving scenarios. By harnessing the power
that each journey is optimized for comfort. of dynamic damping control, this innovation bridges
Enhanced stability: By synchronizing the damping the gap between comfort and performance, creating
characteristics of the front and rear wheels, electric a harmonious synergy that benefits both driver and
shock absorbers enhance vehicle stability. This syn- passengers. Furthermore, the universal application of
chronicity mitigates unwanted oscillations, mini- this technology goes beyond luxury vehicles, extend-
mizing body roll during cornering, and providing a ing its advantages to a broader spectrum of automo-
heightened sense of control to the driver. biles. This democratization of advanced suspension
Optimized performance: Electric shock absorbers systems not only enhances driving experiences but
allow for optimized performance in various driving also democratizes safety and comfort, affirming its
scenarios. Whether navigating through city traffic, role in shaping the future of transportation. As elec-
cruising on highways, or tackling challenging terrains, tric shock absorbers with variable outputs pave the
the damping characteristics can be fine-tuned for opti- way for a new era in vehicular comfort, stability,
mal handling and response. and control, the potential for continued innovation
Responsive handling: The dynamic adjustment of and refinement remains vast. With the convergence
damping ratios results in improved responsiveness of technology, engineering expertise, and a commit-
to driver inputs. Quick adjustments to damping in ment to improving mobility, this concept propels the
response to sudden maneuvers enhance the vehi- automotive industry toward a horizon where every
cle’s agility, making it more predictable and safer to journey is defined by harmony, performance, and an
handle. unparalleled connection between driver, vehicle, and
Decreased vibration: Vibrations transmitted from the road.
the road to the vehicle’s chassis are greatly reduced by
the variable dampening function. The ride is smoother References
and more comfortable for passengers, who are spared Audi Technology Portal: Dynamic Ride Control. Available
the annoyance of uneven and poor roads at: [Link] [Link]/en/chassis/
Adaptive suspension: Electric shock absorbers suspension controlsystems/dynamic- ride-control_en
with variable outputs form the basis for an adaptive [Retrieved 1st August 2018].
Applied Data Science and Smart Systems 333
Shams, M., Ebrahimi, R., Raoufi, A., and Jafari, (2007). a magnetorheological damper. J. Vib. Con., 10(3),
CFD-FEA analysis of hydraulic shock absorber valve 461–471.
behavior. Int. J. Autom. Technol., 8(5), 615–622. Gupta, A. et al. (2006). Design of electromagnetic shock ab-
Motor Trend: 2014 Chevrolet Corvette Stingray Z51 sorbers. Int. J. Mech. Mat. Des., 3(3), 285–291.
First Test. Available at: [Link] Hrovat, D. (1997). Survey of advanced suspension develop-
ca/en/news/2014-chevrolet- corvette-stingray-z51- ments and related optimal control applications. Auto-
firsttest/#2014-chevrolet-corvette- stingray-z51-sus- matica, 33(10), 1781–1817.
pension [Retrieved 1st August 2018]. Hu, Hongsheng, Xuezheng Jiang, Jiong Wang, and Yancheng
Popular Mechanics: 3 Technologies That Are Making Car Li. (2012). Design, modeling, and controlling of a
Suspensions Smarter Than Ever. Available at: large-scale magnetorheological shock absorber under
[Link] tech- high impact load. Journal of Intelligent Material Sys-
nology/a14665/why-car-suspensionsare-better-than- tems and Structures 23(6): 635–645.
ever/ [Retrieved 1st August 2018]. Raman, R. S., Basavaraj, Y., Prakash, A., and Garg, A.
Tiwari, S., Singh, M. K., and Kumar, A. (2020). Regenera- (2017). Analysis of six sigma methodology in export-
tive shock absorber. Res. Rev., 9, 565–569. ing manufacturing organizations and benefits derived:
Thakur, D., Singh, J., Dhiman, G., Shabaz, M., and Gera, A review. 2017 3rd Int. Conf. Comput. Intel. Comm.
T. (2021). Identifying major research areas and minor Technol. (CICT), 1–5.
research themes of android malware analysis and de- Irmscher, S., and E. Hees. (1996). Experience in semi-ac-
tection field using LSA. Complexity, 2021, 1–28. tive damping with state estimators. In proceeding of
Faheem, Ahmad, Fairoz Alam, and Varikan Thomas. AVEC, 96, 193–206.
(2006). The suspension dynamic analysis for a quarter Lee, M., Shamsuzzoha, M., and Luan Vu, T. N. (2008).
car model and half car model. In 3rd BSME-ASME IMC-PID approach: An effective way to get an ana-
International conference on thermal engineering, Dha- lytical design of robust PID controller. Int. Conf. Con.
ka, 20–22. Autom. Sys., 2861–2866.
Fateh, M. M. and Alavi, S. S. (2009). Impedance control Milliken, William F., Douglas L. Milliken, and L. Daniel
of an active suspension system. Mechatronics, 19(1), Metz. (1995). Race car vehicle dynamics. Vol. 400.
134–140. Warrendale: SAE international.
Amr Mansour, S. (2011). DC motor control using ant colo- Tran, Vu-Khanh, Pil-Wan, H., and Yon-Do, C. (2022). De-
ny optimization. sign of a 120 W electromagnetic shock absorber for
Guo, D., Hu, H., and Yi, J. (2004). Neural network motorcycle applications. Appl. Sci., 12(17), 8688.
control for a semi-active vehicle suspension with
44 Integrating metaverse and blockchain for transparent and
secure logistics management
A.U. Nwosu1, S.B. Goyal2,a, Anand Singh Rajawat3, Baharu Bin Kemat4
and Wan Md Afnan Bin Wan Mahmood5
City University, Petaling Jaya, 46100, Malaysia
1,2,4,5
3
School of Computer Science & Engineering, Sandip University, Nashik, Maharastra, India
Abstract
The logistics industry plays an essential role in global commerce by ensuring goods’efficient movement and transportation
across various supply chains. However, as the logistics network keeps expanding due to the emergence of e-commerce, tradi-
tional logistics management systems face challenges related to maintaining transparency and security. Integrating blockchain
with the metaverse can revolutionize and offer solutions to the issues mentioned earlier. Blockchain is an immutable and
decentralized technology disrupting operations in different areas such as healthcare, banking, smart city, and logistics. This
paper aims to address the abovementioned problem in real-life computer interaction scenarios. It highlights the potential
benefits of integrating metaverse in logistics. It proposes a blockchain-based logistics management system with the meta-
verse’s immersive virtual environment to enhance the security and transparency of logistics management systems. The system
efficacy was tested based on privacy, security, latency and throughput. The proposed approach is more secure and efficient
compared with the existing system.
Keywords: Metaverse, blockchain technology, logistics management, transparency, virtual reality, augmentative reality, smart
contract
drsbgoyal@[Link]
a
Applied Data Science and Smart Systems 335
the paper concludes the study and provides a future Interconnectivity: The metaverse comprises inter-
research agenda and scope. connected virtual spaces, often called “worlds” or
“domains.”Individuals, organizations, or communi-
Background ties can create these spaces, which can be linked to-
gether, allowing users to navigate between different
Definition and characteristics of metaverse virtual environments seamlessly.
Metaverse is a collective virtual shared space cre- User-generated content: Users play a crucial role in
ated by converging virtually enhanced physical real- shaping and expanding the metaverse through creat-
ity and physically persistent virtual reality (Khattar ing and sharing content. They can build virtual ob-
et al., 2020; Ritterbusch et al., 2023). It is also an jects, environments, and experiences, contributing to
immersive, interconnected, and interactive virtual the richness and diversity of the virtual universe.
universe where users can engage with each other and
the virtual environment in real time. The Metaverse Real-time interaction: The metaverse enables real-
can be accessed through various devices such as vir- time interaction and communication among users. It
tual reality headsets, augmented reality glasses, com- includes voice and text-based chat, virtual meetings,
puters, and mobile devices. Figure 44.1 depicts the collaborative workspaces, and social interactions,
architecture and layers of the metaverse (Al-Ghaili et fostering a sense of presence and social connection
al.,2022). within the virtual environment.
The following are the characteristics of a metaverse. Cross-platform accessibility: The metaverse aims to
be accessible across different platforms and devices,
Immersion: Metaverse provides a highly immersive ensuring users can engage with the virtual world re-
experience by simulating a three-dimensional envi- gardless of their chosen hardware or operating sys-
ronment where users can navigate and interact. It of- tem.
ten incorporates virtual reality (VR) and augmented
reality (AR) elements to create a sense of presence Opportunities of metaverse in logistics operations
within the virtual world. Metaverse has the capability of optimizing and
enhancing logistics operations. This immersive tech-
Shared space: The metaverse is where multiple users
nology offers unique opportunities to improve effi-
can interact and collaborate. Users can communicate
ciency, training, visualization, and decision-making
with each other, engage in activities, and create con-
processes within the logistics industry. Here are some
tent within the virtual environment.
critical potential metaverse in logistics operations:
Persistence: The metaverse maintains a persistent ex-
istence, meaning it continues to exist and evolve even Training and simulation: Metaverse can be used
when users are not actively present. User changes per- to create realistic and interactive training simula-
sist over time, allowing for the development of a dy- tions for logistics personnel. It includes training for
namic and evolving virtual world. warehouse workers, truck drivers, and other logistics
professionals. By providing a safe and controlled envi- a. Data security and privacy: Logistics manage-
ronment, trainees can practice tasks, learn operational ment involves exchanging sensitive information,
procedures, and improve their skills without needing including shipment details, customer data, and
physical resources or putting valuable goods at risk. financial transactions. Integrating the metaverse
Warehouse management: Metaverse can provide introduces new security risks, as virtual environ-
warehouse workers with real-time information and ments may become vulnerable to hacking, data
guidance. Using AR-enabled smart glasses or devices, breaches, and unauthorized access.
workers can see digital overlays of product loca- b. Interoperability: As the metaverse evolves, vari-
tions, picking instructions, and inventory data, help- ous platforms and technologies emerge, each
ing them navigate the warehouse more efficiently and with its standards and protocols. Ensuring in-
accurately. teroperability between different metaverse sys-
tems and logistics media is critical for seamless
Load planning and cargo visualization: Metaverse
data exchange and collaboration across multiple
can assist load planning by creating 3D virtual rep-
stakeholders.
resentations of cargo and containers. Logistics man-
c. Cost and investment: Developing and imple-
agers can visually inspect how items fit together and
menting a metaverse logistics management sys-
ensure optimal use of available space in containers or
tem can be costly, especially for smaller logistics
trucks, reducing wastage and minimizing the risk of
companies. The expenses associated with hard-
damage during transit.
ware, software, training, and maintenance may
Last-mile delivery optimization: Metaverse can aid present a barrier to entry for some organizations.
delivery drivers in finding the most efficient routes Balancing the potential benefits with the initial
and locating specific delivery addresses. The AR navi- investment is a crucial consideration.
gation can overlay directions onto the driver’s field of d. Scalability: As logistics operations scale up and
view, allowing them to stay focused on the road while more users join the metaverse system, the in-
receiving real-time navigation updates. frastructure needs to accommodate increased
Remote assistance and collaboration: The AR com- demand and maintain a consistent level of per-
ponent of metaverse facilitates remote collaboration formance. Scalability challenges may arise, re-
and assistance for logistics professionals. For exam- quiring continuous monitoring and adjustments
ple, experts can use AR technology to guide on-site to handle growing user loads.
workers through complex repair or maintenance
procedures, reducing downtime and increasing opera- All these challenges can be addressed by leveraging
tional efficiency. blockchain-based solutions in logistics management
Quality control and inspection: Metaverse can be in a metaverse environment.
used for the virtual inspection of goods, especially in
cases where the physical presence of an inspector is Overview of blockchain
challenging or costly. It can improve the accuracy and Blockchain is a disruptive decentralized technol-
speed of quality control processes in logistics. ogy that enables secure, transparent, and immutable
transaction records (Zheng et al., 2017). Each block
Real-time tracking and supply chain visualization:
in the blockchain records information and is linked
Metaverse an immersive view of the entire supply
together using cryptographic techniques (Swan,2017).
chain, allowing logistics managers to monitor ship-
Blockchain has gained significant popularity with
ments, track goods in real-time, and identify potential
the rise of cryptocurrencies, most notably Bitcoin
bottlenecks or delays.
(Leekha,2018). It has been applied beyond digital
Customer experience: In the context of e-commerce currencies like healthcare, smart cities, and supply
and retail logistics, AR can enhance the customer chain management. Figure 44.2 depicts the transac-
experience by enabling virtual try-ons, product visual- tion process of blockchain technology.
ization, and interactive shopping experiences, increas-
ing customer satisfaction and reducing the likelihood Smart contract and logistics automation
of product returns. Smart contracts have the potential to revolutionize
logistics automation by introducing trust, transpar-
Challenges of logistics integrated metaverse ency, and efficiency into various aspects of the supply
Integrating the metaverse and logistics management chain. A smart contract is a self-executing program
system holds great promise, but it also comes with with the terms of the agreement directly written into
several challenges that must be addressed for suc- code (Li et al., 2020). Once the pre-defined condi-
cessful implementation. Some of the key challenges tions are met, the contract automatically executes the
include (Allam et al., 2022) are as follows: specified actions without intermediaries or manual
Applied Data Science and Smart Systems 337
Kamble et al., 2022 Digital twin integration with logistic supply Metaverse was not integrated into the system
chain
Subramanian et al., 2020 Blockchain smart contract-based fourth party No integration of metaverse
logistics
Paliwal et al., 2020 Blockchain-based robust framework No integration of metaverse environment
Tan et al., 2023 Metaverse-based logistics and marketing This work did not integrate blockchain
Roy et al., 2023 Metaverse based on teaching This work does not focus on logistics
management
Ali et al., 2023 Blockchain-based and AI metaverse in the This work does not focus on the logistics
healthcare system system
inquiry about the condition and location of goods Table 44.2 Tools and specifications
is done. Here, the goods also have their avatar. The
[Link]. Tools Specifications
manufacturer can request the state and condition of
the warehouse, and the transporter can have a video 1 Remix IDE Intel (R) core of i5
monitoring of the goods while on the road. All data 8250U
created during the transaction is stored in the block- 2 Solidity 0.7.0
chain repository.
3 Decentral and explore Intel HD/UHD 9th gen
Algorithms of the proposed system 4 Window 11, personal 1.6 GHz, 8 GB of RAM,
computer with a 64-bit operating
Algorithm 1: Activation of customer environment and system
customer inquiry
Input: Customer, Transporter, Manufacturer,
Output: Activation of Metaverse Environment and
with the physical environment. Table 44.2 depicts the
Initiation of Customer Logistics Inquiries
specifications of deployed tools.
1: Procedure: Blockchain_LogisticMeta ()
2: if (C_ID== True) then Results and analysis
3: Display the avatar of the Customer
This section presents the proposed system’s simula-
4: else
tion results and the proposed system’s evaluation of
5: Display customer does not exist
the existing system. Figure 44.4 depicts the simulation
6: if (CP= true), then
results of the proposed approach.
7: Execute _contract (for the Customer)
8: Setup Customer Virtual Environment
Privacy and security evaluation
9: else
The privacy and security evaluation of the proposed
10: return to none
system is tested based on the following cyber threats:
10: end if
11: end Insider attack: This attack occurs when a logistics user
Note: CP= Customer private key accesses private data or information without legiti-
macy. The system protects against this attack, which
hashes the data while transmitted along the network.
Algorithm 1 shows the flow of information in both
blockchain and logistics metaverse. When a customer DDoS attack: This attack occurs when an adversary
inquires about the status of his goods from the logis- floods the system network with malicious code to
tics company, the system first conforms to the authen- shut and breach the communication channel. The
ticity of the customer. Customers who need to register proposed system mitigates this attack using a decen-
will be redirected to register with the system. The cus- tralized node and consensus mechanism.
tomer smart contract can only be executed when the One-point-of-failure attack: The attack happens
private key is correct, and the interaction of custom- when the system gets corrupted or compromised by
ers with the logistics company will be in the virtual introducing a corrupted device, halting the whole
environment. The customer can track the location of system. The system protects against this attack using
goods virtually. Once the inquiry is completed, it will decentralized nodes and device authentication.
be stored in the blockchain, and the environment will
disappear. Performance evaluation
Furthermore, other logistics stakeholderslike trans- The performance evaluation of the proposed sys-
porters and manufacturers follow the same pattern to tem was tested based on two metrics: latency and
activate their environment. throughput.
Experiment environment setup Latency: This refers to the time frame between trans-
The experiment aims to implement the proposed sys- action initiation and transaction completion time.
tem (blockchain-based metaverse integrated logistics Table 44.3 depicts the latency result of the registra-
management system) that will enhance real-time data tion process between the proposed and existing logis-
sharing and security in logistics operations. The smart tics systems.
contracts are created using the solidity version and
deployed using Remix IDE. The decentral platform Figure 44.5 shows the analysis of latency results.
makes a virtual representation of the logistics system. It shows that the proposed system has higher latency
Web 3.0 is used to connect the virtual environments than the existing baseline system.
340 Integrating metaverse and blockchain for transparent and secure logistics management
Table 44.3 Latency results between the proposed system Table 44.4 Throughput comparison between the baseline
and the existing system system and the proposed system
Abstract
The research aims to identify the algorithms and techniques that have been applied to the identification of heart disease.
Since there are more and more occurrences of heart disease every day, it is important and difficult to anticipate any
prospective problems. This diagnosis is a difficult task that demands precision and effectiveness. The early detection
of cardiovascular diseases depends on heart sound analysis. Practically speaking, the advancement of computer-based
heart sound analysis is appealing. This paper is the survey of different algorithms and approaches that can be used to
find heart disease and there can be various attributes for the same like speed, accuracy. The suggested method focuses
on automatically classifying phonocardiogram (PCG) data after removing noise using a convolution neural network in
order to lessen the need on skilled medical professionals for heart sound detection. Algorithms that are compared in this
paper are support vector machine (SVM), convolutional neural network (CNN) with and without augmentation. Because
of their ability to analyze images accurately, CNN have quickly attracted the interest of researchers and medical profes-
sionals. In order to diagnose cardiovascular disease, this study sought to design a system that combines various machine
learning techniques, such as K-nearest Neighbor, Naive Byes, linear regression, decision tree, Alex-Net, ensemble learning
and random forest.
Hence this paper gives a relative study of numerous approaches that were used to classify and detect cardiac diseases.
Keywords: Cardiac diseases, SVM, Naïve-Bayes, Alex-Net, CNN with and without augmentation
a
dhowmyab@[Link]
344 A systematic study of multiple cardiac diseases by using algorithms
cost of medical care and the standard of care given to These PCG signals are shown visually in five dif-
patients may be impacted by this. Therefore, we must ferent forms in the following example in Figure 45.4.
create a productive system for reducing human mis- The rest of the article is organized as follows: The
take and raising patient care quality. This is possible by related work in the field of cardiac diseases. Followed
fusing computer decision support systems with medi- by conclusion and future work.
cal decision support systems (Figures 45.2 and 45.3).
Cardiac diseases are serious and need to be accu- Related work in the field of cardiac diseases
rately recognized at an early stage utilizing routine
auscultation tests. Heart auscultation is a crucial In this section we define the objective and technique
component of a heart examination used in medicine used, and accuracy achieved.
to detect early-stage cardiac disorders. Cardiac aus- In paper by Baghel et al. (2020), convolutional neu-
cultation is a technique for listening and analyzing ral network also knows as CNN model is used in the
to heart sounds (Baghel et al., 2020). Human cardiac proposed system because of its excellent accuracy and
auscultations are examined with a stethoscope. A robustness in autonomously diagnosing cardiac dis-
traditional stethoscope is used in clinical settings to eases from heart sounds. In order to improve precision
examine the health of a human heart. It is a simple, in a noisy environment and make the system resilient.
effective method that also costs nothing computation- For multi-classification and training 2124qof various
ally, but understanding and interpreting heart sounds cardiac conditions, the proposed method has utilized
requires medical training (Leatham, 1975). data augmentation techniques.
Clinical interpretation of the cardiac auscultations Results of this paper are – N-fold cross-valida-
may only be done by a qualified medical specialist. tion and both heart sound data were used to vali-
We will use ML-based automatic classification system date the model with enriched data. All fold’s results
based on heart sounds to diagnose cardiac disorders. have been displayed and published in this work. The
Computerized heart sound recording is known as model utilized in this study had a 98.60% accuracy
a phonocardiogram, or PCG. phonocardiography rate on tests designed to identify numerous heart
(PCG). PCG is a non-invasive, cost-effective method disorders
of recording heart impulses. In paper by Tang et al. (2018), the support vector
A number of cardiovascular disease (CVD) sig- machine (SVM) classifier’s powerful classification
nals, such as mitral stenosis (MS) aortic stenosis (AS) ability is demonstrated by the characteristics. The
mitral regurgitation (MR) and mitral valve prolapse outcomes demonstrate that the overall score which
(MVP), can be diagnosed using PCG signals. was determined by 200 independent simulations,
Figure 45.4 Utilizing a phonogram (Baghel et al., 2020) signals from the current CVD classes. (a) Aortic stenosis (AS),
(b) Mitral regurgitation (MR), (c) Mitral stenosis (MS), (d) Mitral valve prolapse (MVP) and (e) Normal
346 A systematic study of multiple cardiac diseases by using algorithms
is 0.880.02, which is comparable to the perfor- study combined with a traditional feature engineering
mance of the previous top classification techniques. technique. In the beginning, 497 characteristics were
Furthermore, the SVM classifier performs admirably yielded by 8 domains.
with even a minimal number of training features and To obtain global information about features and
consistently produces reliable results with randomly avoid over fitting. The features are embedded into the
chosen training features. Five hundred and fifteen built-in CNN which is usually used before the clas-
features used in this study are time interval, state sification layer but excludes all connected processes
amplitude, energy, high-order statistics, cepstrum, fre- involving the global average layer.
quency spectrum, cyclostationarity, and entropy. The To enhance the effectiveness of the classification
frequency spectrum features have the greatest classifi- method, the class weights of the loss function were
cation-contributing value, according to a correlation set in the training phase while accounting for the
analysis between the features and the target label. class imbalance. The effectiveness of the suggested
Haya Alaskar et al. (2019) and Gera et al. (2021) technique was assessed using stratified 5-fold cross-
data gathered from the 2016 PhysioNet/CinC chal- validation. Matthews correlation coefficient, mean
lenge dataset. In these papers, it examines the perfor- accuracy, sensitivity, and specificity, respectively,
mance of a CNN named AlexNet, concentrating on 72.1%, 86.8%, 87%, and 86.6% on the PhysioNet/
two methods for identifying abnormalities in PCG CinC challenge 2016 dataset. The proposed method
signals. Heart sound recordings from both clinical performs well in terms of sensitivity and specificity.
and non-clinical settings are included in this dataset. In a paper by Abbas et al. (2022), using a unique
Our thorough simulation findings showed that 87% attention-based technique, CVT-Trans also known
recognition accuracy was reached utilizing AlexNet as as convolutional vision transformer recognizes and
the feature extractor and SVM as the classifier. This is classifies PCG signals majorly in five groups. The
an improvement of 85% accuracy attained by end-to- CWTS also known as continuous wavelet transform-
end learning AlexNet in contrast to the benchmarked based spectrogram method was used to extract fea-
methodologies. tures from the PCG data. Accuracy of 100%, SE of
In a work done by Li et al. (2020), cardiac diseases 99.00%, SP of 99.5%, and F1-score of 98% indicate
are diagnosed using heart sound as a key component. the overall average accuracy.
Experts struggle and take a lot of time to distinguish In a work did by Son et al. (2018) explains that
between various heart sounds because of the low sig- 97% accuracy rate is achieved for diagnosing patients
nal-to-noise ratio (SNR). The scientific classification with heart problems. Mel frequency cepstral coeffi-
of heart sounds is therefore necessary. To automati- cient, discrete wavelets transform, and centroid dis-
cally distinguish between normal and diseased heart placement-based k-nearest neighbor features of the
sounds, we used deep learning algorithms in this heart sound signal were used to extract features, while
[1] Cardiac disorders detection using multi- CNN with augmentation Accuracy achieved – 96.23%
classification algorithm CNN without augmentation Accuracy achieved – 98.60%
[2] Binary classification between abnormal SVM Accuracy – 88%
sounds and normal PCG signals
[3] Binary classification between abnormal AlexNet, SVM Accuracy – 87.65%
sounds and normal PCG signals
[4] Using PCG signals classification between Ensemble of a feature Accuracy – 86.02%
normal and abnormal sounds engineering method, deep
learning algorithm
[5] Cardiac disorders detection using multi- Deep learning Accuracy – 100%
classification algorithm SE – 99.00%
SP – 99.5%
F1-score – 98%, based on
10-fold cross validation
[6] Heart diseases detection using multi- MFCC’c Accuracy – 97%
classification algorithm DWT (using SVM and
DWT)
[7] Several heart diseases detection using a multi- SVM, K-NN Accuracy – 97.78%
classification algorithm
Applied Data Science and Smart Systems 347
SVM, deep neural network (DNN), and centroid Med., 197. Epub 105750. [Link]
displacement-based k-nearest neighbor were utilized cmpb.2020.105750.
for learning and classification. For training and clas- Tang, Hong, Ziyin Dai, Yuanlin Jiang, Ting Li, and Chengyu
sification using SVM and discrete wavelets transform Liu. (2018). PCG classification using multidomain fea-
tures and SVM classifier. BioMed research internation-
(DWT), we merged Mel frequency cepstral coefficient
al, Vol. 2018. [Link]
and DWT features. This improved the results and
Alaskar, H., Alzhrani, N., Hussain, A., and Almarshed, F.
classification accuracy. Results can be significantly (2019). The implementation of pretrained AlexNet on
enhanced when Mel frequency cepstral coefficient and PCG classification. Int. Conf. Intel. Comput., 784–794.
DWT features are combined and used for classifica- [Link]
tion via SVM, DNN, and k-nearest neighbor (KNN). Li, F., Tang, H., Shang, S., Mathiak, K., Cong, F. Clas-
In a paper by Yadav et al. (2019), uses heart sounds sification of heart sounds using convolutional neu-
as the input. The suggested system uses frame-based ral network. Appl. Sci., 10(11), 3956. [Link]
processing and strategic processing to extract ML org/10.3390/app10113956.
features that can distinguish between different heart Abbas, Q., Hussain, A., and Baig, A. R. Automatic detec-
sounds. A supervised classifier is trained to automati- tion and classification of cardiovascular disorders
using phonocardiogram and convolutional vision
cally detect heart problems using the most significant
transformers. Diagnostics, 12(12), 3109. [Link]
features. Differences in auscultations are brought on
org/10.3390/diagnostics12123109.
by biological anomalies that affect how the heart Gera, T., Singh, J., Mehbodniya, A., Webber, J. L., Shabaz,
physically works. for the classification of abnormal M., Thakur, D. (2021). Dominant feature selection
and normal heart sounds, the suggested method had and machine learning-based hybrid approach to ana-
an accuracy of 97.78% and an error rate of 2.22% lyze android ransomware. Sec. Comm. Netw., 2021,
(Table 45.1). 1–22. [Link]
Son, G.-Y., and Kwon, S. (2018). Classification of heart
sound signal using multiple features. Appl. Sci., 8(12),
Conclusion and future scope 2344. [Link]
In this paper, the main emphasis is given to the Yadav, A., Singh, A., Dutta, M. K., and Travieso, C. M.
involvement of several research works available in a (2019). Machine learning-based classification of
digital repository like IEEE, springer for heart disease cardiac diseases from PCG recorded heart sounds.
Neu. Comput. Appl., 1–14. [Link]
detection from 2016 to 2022 onwards. The system-
s00521-019-04547-5.
atic study clearly shows the attainments being done
World Health Organization. Cardiovascular diseases.
in heart disease detection with proper accuracy rates [Link]
from the past many years. cardiovascular-diseases- (cvds)#.[Link],
ML algorithms generally produced encouraging May 2017. Accessed: Oct 2019.
results, despite the fact that there are still a number Upretee, P. and Yuksel, M. E. (2019). Accurate classifica-
of obstacles to be cleared before they can be used in tion of heart sounds for disease diagnosis by a sin-
clinical practice. To interpret the study in the appro- gle time-varying spectral feature: Preliminary re-
priate clinical context, it is necessary to choose the sults. 2019 Sci. Meet. Elec.-Electron. Biomed. Engg.
suitable algorithms for the relevant research ques- Comp. Sci. (EBBT), 1–4. [Link]
tions, compare the results to those of human special- EBBT.2019.8741730.
Fu, W., Yang, X., and Wang, Y. (2010). Heart sound diag-
ists, use validation cohorts, and report on all potential
nosis based on DTW and MFCC. 2010 3rd Int. Cong.
assessment matrices. Most significantly, investigations
Image Sig. Proc. (CISP), 2920–2923. [Link]
comparing ML algorithms to traditional risk models org/10.1109/cisp.2010.5646678.
should be conducted in the future. Once validated in Gudadhe, M., Wankhade, K., and Dongre, S. (2010). Deci-
this manner, ML algorithms could be implemented in sion support system for heart disease based on support
clinical settings and integrated with electronic health vector machine and artificial neural network. 2010
record systems, especially in regions with abundant Int. Conf. Comp. Comm. Technol. (ICCCT), 741–745.
resources. [Link] 5640377.
In the future, we’ll work to identify different heart Shashikant, G., P. Chetan, and G. Ashok. (2011). Heart
disease subtypes and further classify each one accord- Disease Diagnosis using Support Vector Machine. In
ing to how severe it is. International Conference on Computer Science and
Information Technology.
Sheela, C. J. and Vanitha, L. (2014). Prediction of sudden
References cardiac death using support vector machine. 2014 Int.
Conf. Cir. Power Comput. Technol. (ICCPCT), 377–
Baghel, N., Dutta, M. K., and Burget, R. (2020). Automatic
381. [Link]
diagnosis of multiple cardiac diseases from PCG sig-
Leatham, A. (1970). Auscultation of the heart and phono-
nals using convolutional neural network. Nat. Lib.
cardiography. Churchill Livingstone, London.
46 Forecasting mobile prices: Harnessing the power of
machine learning algorithms
Parveen Badoni1,a, Rahul Kumar2, Parvez Rahi3, Ajay Pal Singh Yadav4 and
Siroj Kumar Singh5
Department of CSE, Chandigarh University Mohali, Punjab, India
1,2,3,4
5
Department of CSE, HMRITM, Hamidpur, New Delhi, India
Abstract
The primary objective of this paper is to forecast optimal prices for top-tier smartphones, while considering their available
features. Our approach involves the development of a machine learning (ML)-based price range prediction model, which
harnesses various algorithmic techniques applied to an extensive dataset. This model generates comprehensive data visual-
ization, aiding decision-making processes. Additionally, our proposed model facilitates market analysis within the sector
by comparing its accuracy against other models. For instance, numerous companies engaged in the purchase of pre-owned
mobile phones employ their proprietary models. Users can cross-reference these models with ours to pinpoint the most suit-
able price for their mobile devices. The accuracy level can be gauged against alternative models to obtain the most reliable
results. Our research encompasses a wide array of features and events to predict mobile phone prices, thereby addressing
buyer concerns and simplifying their quest for smartphones within their budgetary constraints.
Keywords: Machine learning, mobile prices, linear regression, predictive model, KNN, SVM, smart phone prices
rir7890@[Link]
a
Applied Data Science and Smart Systems 349
highly accurate predictions due to the influence of known input and output data, enabling them to pre-
these unpredictable factors. dict future outputs. Unsupervised learning uncovers
Key considerations in the modeling process encom- hidden patterns within the input data, while super-
pass data collection, the identification of significant vised learning identifies and leverages patterns already
features, and a comprehensive analysis of both recent present in the data. Both the data and the computa-
and historical developments within the mobile sector. tional complexity can be decreased through feature
Factors such as the brand’s reputation, economic con- selection. It can also become more effective and iden-
ditions, competitor analysis, regression analysis, and tify the feature subsets that are valuable (Thu Zar and
the application of ML techniques all play pivotal roles Nyein, 2016).
in constructing effective predictive models. These fac- In this research paper, various supervised and unsu-
tors collectively contribute to a more holistic under- pervised learning methods are employed to predict
standing of mobile pricing dynamics. The bulk of this our model. The data is carefully prepared to enhance
research paper is dedicated to the implementation of precision in our model. To collect data for our model,
a judicious selection of variables in mobile valuation specific websites and links are utilized, simplifying the
techniques. This process is instrumental in identifying data acquisition process and allowing for the accumu-
which variables are the most pertinent and appropri- lation of a substantial dataset. The epsilon, polyno-
ate to include in the model. The knowledge acquired mial degree, and gamma are the three most significant
through this research has broader implications, optimum parameter values that evolution strategy
enabling various fields to gain insights into the cir- can converge to more quickly than cost (Listiani et
cumstances that warrant specific studies and the occa- al., 2009).
sions when suitable techniques should be applied. In Within the paper, a variety of ML algorithms are
this dynamic environment, statistical models are more applied, including K-nearest neighbors (KNN), sup-
suited for short-term forecasting because they by defi- port vector machine (SVM), support vector regression
nition reflect actual market results (McMenamin and (SVR), as well as linear and polynomial models. These
Monforte, 2000). diverse algorithms are harnessed to predict the out-
In this context, the primary challenge lies in pre- put within our model. The inclusion of multiple algo-
dicting our model based on both market prices and rithms serves the purpose of enhancing the accuracy
the key features of mobile devices. This is achieved of our model, ensuring a comprehensive approach to
through the utilization of the support vector machine mobile price prediction.
(SVM) concept. Previous research has indicated that,
especially when dealing with large datasets, the SVM data collection
technique outperforms other methods, such as mul-
tiple linear regression, in terms of accuracy for price The dataset includes mobile phones manufactured or
prediction. Using back propagation algorithms to assembled by various companies, including Samsung,
simulate and predict runoff forecasting results and Apple, Google, BBK Electronics Corporation, and
comparing expected forecasting result accuracy with others. Interestingly, whether a mobile phone has a
existing forecasting approaches may ultimately result memory card slot or not is considered a notable fea-
in a more trustworthy data mining strategy (Mishra ture, highlighting the importance of this aspect in the
et al., 2014). dataset.
The central aspect of our mobile prediction model Several features in the dataset have numerical val-
revolves around the unique approach of predicting ues, including display size, thickness, internal memory
mobile prices based on their model names and key size, camera pixel size, RAM size, and battery size.
features, which are available on the internet. This rep- These numerical values offer insights into the specifi-
resents a pivotal distinction, positioning our model cations of the mobile phones.
a step ahead of other predictive mobile models. Figure 46.1 provides a description of the dataset,
While various types of models have been employed including statistical measures such as mean, median, and
for mobile price prediction, many of them fall short count. This information is crucial because it helps assess
in this critical aspect. Although one method that the characteristics and distribution of the data. The bal-
improves prediction performance is data cleansing, ance of the dataset, in terms of these statistical measures,
it is insufficient when dealing with complicated data is used to evaluate the fitness of the data for analysis
sets like the one used in this study (Gegic et al., 2019). and modeling. Ensuring that the data is balanced and
representative is essential in selecting an appropriate
dataset for research, as it can significantly impact the
Methodology
quality and reliability of the results. Therefore, under-
Machine learning employs both supervised and unsu- standing the data’s description and balance is vital for
pervised learning approaches to train models using the data selection process in this research.
350 Forecasting mobile prices: Harnessing the power of machine learning algorithms
In our research, we’re faced with the challenge of per the price in the online apps. We are using it in our
classifying mobile devices as either very pricey or daily life, which we can use the same in our daily life.
not. However, the continuous and dynamic nature of Now let’s highlight about description or how to take
mobile device prices in our rapidly changing society data. The price of the data here is taken from amazon,
complicates this task. This led us to transform the eBay, Flipkart, etc., are our source of the data that we
regression problem into a classification model. In collected to predict price as per festivals offers and
this classification, we’ve grouped the mobile device normal discounts are given by the apps or by the com-
prices into four classes, although prices are con- panies. All the features are the same, but we added
tinually changing. Both decision trees and the naive more columns predicting our mobile price.
Bayes classifier, however, have a fundamental limita-
tion when it comes to handling output values repre- Dimensionality reducation
sented as classes with numerical values. Therefore,
we had to discretize the pricing attributes into these By acquiring a set of key variables, or features, we
classes, which encompass a range of prices. This dis- design a prediction model by limiting the amount of
cretization introduces new potential sources of error random variables that are taken into consideration.
into our model and other related processes. To eval- Data opening fix technique for important data in the
uate the effectiveness of our model, we have split request to create the partitions required to contain
the data into a training set and a test set. The train- the disaster and provide the missing data (Brar et al.,
ing set is used to train the model, while the test set 2022; Nadeem et al., 2023).
is held separate and not used during training. This Prediction model used in this paper is not totally
approach allows us to assess how well the model practical since it becomes more difficult to present
performs on unseen data, providing a measure of the training set and use that dataset to generate pre-
its generalization and predictive power. A customer dictions the more features there are. Most of these
can be recommended a good product by providing functions can occasionally be redundant because they
an economic range (Muhammad and Khan, 2018) are connected with one another, which can reduce
(Figure 46.2). the model’s accuracy. With the help of dimensionality
Many elements, like memory, display, battery life, reduction algorithms in this situation, ML employs
camera quality, and so forth, are taken into account two distinct types of dimensionality reduction tech-
while buying mobile phones. Due to the lack of niques such as feature selection and feature extrac-
resources required to cross-validate the price, peo- tion. During element selection, we are looking for
ple make incorrect decisions (Singh et al., 2019; the dimensions that provide greatest data and filter
Krishnamurthy et al., 2021). The data here is impor- unnecessary data for the model’s prediction.
tant because it describes model perfectly. Now the Using feature extraction, main goal is to identify a
main question arises why we chose this format to new set of dimensions getting that result from com-
represent data? Simple answer is that it gives data as bining the original dimensions. In machine learning,
Applied Data Science and Smart Systems 351
we typically add as many features to collect key Here are some common data representation meth-
details and produce more accurate results. As the ods and their purposes:
number of elements rises, the model’s performance Heat maps are useful for visualizing the relation-
will eventually start to suffer. This frequently referred ships between data points in a matrix. They are often
as the dimension curse. The problem with dimension- employed to display correlation matrices, making it
ality is that, for our prediction model, sample density easier to identify patterns and dependencies between
falls off exponentially as dimensionality rises. The variables.
dimension of the feature space increases and becomes Correlation heat maps specifically focus on depict-
sparser as we continue to add features while main- ing the correlation coefficients between different
taining the number of training examples. Finding a variables. They help highlight which features are
correct answer for a ML model is significantly simpler strongly correlated or inversely correlated with each
as a result of this rarity, which is very likely to cause other.
any model to over fit (Figure 46.3). Count plots are typically used for categorical data
and help visualize the distribution of categories within
Exploratory data analysis (EDA) a variable. They provide insights into the frequency of
each category.
EDA is a best approach for data analysis using visual The describe() function is a statistical tool that
techniques. It is used to discover the trends, patterns provides summary statistics about the dataset. It cal-
in the data set and to check assumptions using sta- culates measures like mean, standard deviation, mini-
tistical summary and graphical representations of the mum, maximum, quartiles, etc. This function can help
data set. We load the dataset using the Pandas module identify outliers, central tendencies, and the overall
from the python and then print the first five rows. distribution of numerical features.
We use the head() function to print the first five lines Using these representation methods collectively
(Figure 46.4). allows data scientists and analysts to explore and
Different datasets can exhibit various characteris- understand the dataset thoroughly. It aids in identify-
tics, and the choice of data representation methods ing potential data issues, such as missing values or
can significantly impact our understanding of the data outliers, and can reveal valuable insights about fea-
and the modeling process. Various data visualization ture relationships and data distributions. This under-
techniques can be employed to gain insights into the standing is crucial for building accurate predictive
dataset’s features and values, ultimately facilitating models and making informed decisions based on the
the model prediction and elucidating the relation- data.
ships between features. Under the current situation, Addressing missing data in a dataset is a critical
the system uses a technique where a seller randomly consideration in data analysis and modeling. Missing
determines a price without the buyer knowing the data can occur when certain information is not pro-
product’s worth (Balaji et al., 2023). vided for one or more data points within the dataset.
There are various reasons why data may be miss- generates multiple complete datasets with imputed
ing, such as survey respondents choosing not to dis- values and combines results to provide more robust
close certain information (e.g., income or address). estimates.
In such cases, many data values may be absent from Handling missing data is an essential step in data
the dataset, creating a challenge for data analysis preprocessing to ensure that analyses and models are
(Figure 46.5). based on as much available information as possible,
Missing data is a common and real-world problem without introducing bias or inaccuracies due to miss-
in data science and statistics. It can lead to biased or ing values.
inaccurate results if not handled properly. Data scien- In Figure 46.6, a boxplot representation of the
tists and analysts employ various techniques to man- previously mentioned dataset is presented. This
age missing data, such as: graphical representation effectively displays outliers
Imputation: This involves filling in missing values within the dataset. Specifically, it provides insights
with estimated or imputed values based on statisti- into RAM, device width, and device height, which
cal methods. Common imputation techniques include are the primary contributors to changes in the outli-
mean imputation, median imputation, mode imputa- ers graph. Understanding these outliers is crucial as
tion, or using predictive models to estimate missing they can have a significant impact on model predic-
values. tion, especially when their values exhibit substantial
Data collection improvement: In some cases, variations.
improving data collection processes can help reduce To delve deeper into the relationships between
the occurrence of missing data. This may involve bet- various features and their correlation with prices,
ter survey design, clearer instructions to respondents, further analysis is essential for model prediction. The
or data validation checks during data entry. Matplotlib library is employed to showcase these fea-
Deletion: In certain situations, it may be appropri- ture relationships using a Heatmap graph. Heatmaps
ate to remove data points with missing values from are valuable tools for visualizing dependencies and
the analysis. However, this should be done carefully, correlations between different attributes and our
as it can lead to a loss of valuable information and target prediction in the model. This aids in gaining
potential bias. a comprehensive understanding of how various fac-
Advanced imputation: Advanced techniques, such tors influence the pricing of mobile phones. India is
as multiple imputation, can be employed when the expanding, and so is the country’s mobile customer
missing data pattern is complex. Multiple imputation base. India is home to around 900 million mobile
Applied Data Science and Smart Systems 353
Figure 46.7 Heatmap for comparing and choosing the attribute to classify our model
354 Forecasting mobile prices: Harnessing the power of machine learning algorithms
Figure 46.8 Count plot is used to count the occurrence of the observation
phone subscribers, which gives each mobile phone phones excel at this task, offering seamless ways to
manufacturer a stronger platform (Deepesh Kumar, share photos directly online. In the digital age, there
2012). are numerous methods and platforms that facilitate
The corrosion products were analyzed using this process. The megapixels of a cell phone camera
energy-dispersive spectroscopy, scanning electron hold significant importance in the world of photog-
microscopy, and X-ray diffraction (Sidhu, Goyal, and raphy. A higher pixel count directly correlates with
Goyal, 2017) (Figure 46.7). the quality of your photos. Therefore, when seeking
Representation of data in terms of the graph is to capture high-quality images, it’s advisable to con-
given as follows: sider mobile phone cameras with a minimum of 3
megapixels.
(a) Rating/count In essence, the proliferation of digital technology
Rating which is used to rate the mobile phone, it is has made sharing photos online a straightforward
used to get the feedback of the user, who buy the and accessible endeavor, while paying attention to
mobile phone. Its value is in between 0 to 100. This camera megapixels remains a key factor for achieving
graph shows the value, which can easy to understand superior photo quality
(Figure 46.8). The more the pixel of the camera, the more the
price of the mobile increases; but not for the all-
(b) Primary camera/count mobile models it varies from one mobile to another.
Sharing images with the world through social media So this feature is considered in the dataset and
has become a ubiquitous practice when it comes to above is the graphical representation of feature
utilizing a photo gallery. Both iPhone and Android with price.
Applied Data Science and Smart Systems 355
Figure 46.9 Count plot is used to count the occurrence of the RAM
356 Forecasting mobile prices: Harnessing the power of machine learning algorithms
Figure 46.10 Count plot is used to count the occurrence of the primary camera
Figure 46.11 This pairplot is used to understand the best set of feature to explain a relationship between two or more
variables to perform cluster separation
International usability: A smartphone’s ability to Supply chain issues: Global commodity inflation
be used internationally, including factors like network resulting from supply chain disruptions can lead to
compatibility and unlocked status, can influence its increased production costs for mobile phone manu-
price. Devices that offer global usability tend to have facturers. These additional costs may be passed on to
higher price tags. consumers, impacting smartphone prices.
Applied Data Science and Smart Systems 357
Energy prices: Fluctuations in energy prices, par- tweaking by restricting the number of splits and the
ticularly for manufacturing and transportation, can size of nodes can result in gains (Mark, 2004). We
impact the overall cost of producing and distributing Separate our target column by the help of the panda
smartphones. in python and to transform our data in array form
Labor shortages: Labor shortages in manufacturing we used sklearn library to standardized our model
regions can affect production costs, potentially lead- in which we will give this data to our model in that
ing to price adjustments. form for prediction effectively in our price prediction
Input costs: The cost of materials and components model for mobile devices.
required to build smartphones can fluctuate based on
factors like demand, availability, and global economic Training the model
conditions.
Market competition: The competitive landscape In our dataset, which excludes the company column,
within the mobile phone industry can also influence we are employing three primary algorithms to train
pricing. Manufacturers may adjust prices to gain mar- our predictive model: KNN, decision tree, and logistic
ket share or differentiate their products. regression.
It’s important to note that some smartphones may First, let’s delve into how the decision tree algo-
indeed have significantly higher prices compared to rithm is used to predict the model:
others due to factors like advanced features, premium Decision trees are visual representations of deci-
materials, and branding. Understanding these factors sion processes, often depicted as flowcharts, to plan
is essential for consumers looking to make informed and illustrate business and operational decisions. In
choices when purchasing a smartphone. the context of ML, decision trees serve as algorithms
Apple: The important reason is that apple uses to differentiate dataset features using a cost func-
extremely high-quality materials to make their phones tion. The decision tree is initially expanded, and then
robust, which ensure a longer lifespan than android irrelevant branches are pruned through optimization.
phones. For Apple to be able to afford these great mate- Parameters such as the depth of the decision tree can
rials, they need to charge a decent price. iPhones often be adjusted to mitigate the risk of over fitting and cre-
cost more than $1,000, especially for new models. ating overly complex trees.
Samsung: There are many factors, which are listed To implement the decision tree classifier, we start
below. by importing the “[Link]” module in our
Jupyter notebook. Then, we call the “DecisionTree
i. Cheaper build materials – Samsung has recently Classifier()” function to create an instance of the clas-
started adding glass backs to some of its A-se- sifier. For verification purposes, we can print the data
ries phones, but they still generally use lower- to check if it is in the appropriate array form and free
quality metal alloys and finishes than higher-end of errors. Once this is confirmed, we proceed with the
phones. prediction model for mobile prices.
ii. Weaker guts – The higher-end series always uses After creating the classifier instance, we use the
top-of-the-line specifications, while the A-series “fit()” function to train the model using the training
uses the latest mid-range processors with lower dataset and subsequently employ the “predict()” func-
performance but still optimized for battery con- tion to make predictions based on the test dataset.
sumption. This step-by-step process helps us utilize the decision
iii. Weaker camera – Weaker camera sensors are tree algorithm (Table 46.1).
used and less computational photography.
Second KNN used to predict the model
Each data values have some variation, which dis- The KNN algorithm can be contrasted with the most
tinguishes our model by a great margin. That is the precise models since it offers high accuracy for spe-
beauty of the pair plot graph. So, we used Pairplot cific models. Consequently, it finds utility in applica-
by simply called Seaborn library and use Pairplot tions that demand heightened precision without the
function. need for a human-readable model. The predictive per-
formance relies on the distance measure and is par-
ticularly effective for pattern recognition tasks. As a
Preparing data for the model
supervised learning algorithm, it is applicable to both
From here onwards we started our data to prepare, regression and classification problems, although it is
determining what we needed and what we did not. It commonly employed for classification tasks in the
depends on the features what we are choosing in our realm of ML.
model target. Price is our target and the rest of the Neighbors library files are used to predict the
data is used to test and train the model. Additional model. It is the same as the tree classifier, but there is
358 Forecasting mobile prices: Harnessing the power of machine learning algorithms
a difference in using the function to predict the model While logistic regression exhibits similarities with
(Figure 46.12). linear regression, it’s essential to recognize that they
are employed for distinct purposes. Linear regression
Third logistic regression is used here to predict the is employed to address regression problems, where the
model goal is to predict a continuous numerical value, while
Logistic regression is undeniably one of the most logistic regression is specifically tailored for classifica-
widely utilized ML algorithms, particularly in the tion tasks, where the aim is to categorize data into
realm of supervised learning. Its primary objective is predefined classes or groups (Philipp, Wright, and
to predict a categorical dependent variable based on a Boulesteix, 2018).
given set of independent variables. Unlike algorithms
that yield continuous outcomes, logistic regression Data with features and company columns
produces categorical or discrete results, such as “Yes”
or “No,” “0” or “1,” “true” or “false,” and so on. Tree classifier
However, instead of providing precise values like 0 or Supervised learning is a domain where data points are
1, it generates probability values that range between systematically organized based on their predefined
0 and 1. values, aligning with the problem the system aims to
solve. Within this context, decision tree algorithms
prove to be highly efficient and straightforward.
Table 46.1 Summary of parameters
They are often referred to as CART, which stands for
Summary “Classification and Regression Trees.”
Each decision tree comprises a root node from
Correctly Classified Instances 20 which branches extend based on specific conditions,
71.429% leading to leaf nodes. These internal nodes within the
Incorrectly Classified Instances 8 tree represent various test cases applied to the dataset.
28.571% Decision trees are versatile in that they can effectively
Kappa Statistic address both classification and regression problems.
0.6177 These techniques find widespread application across
Mean Absolute Error various industries, providing practical solutions to
0.2066 everyday challenges.
Root mean squared error An apt analogy for this algorithm is a tree structure
0.3668 where predictions are made through the use of dis-
Relative absolute error tinct branch parameters that have been finely tuned
54.2608% (Seifert, Gundlach, and Szymczak, 2019).
Root relative squared error
81.8652% Random forest
Total Number of Instances 28 The random forest algorithm is a versatile supervised
ML method applicable to both classification and
regression tasks. It employs the concept of ensemble
learning, where multiple classifiers collaborate to
address intricate problems and enhance model per-
formance. In our context, integrating random forest
into our price prediction model is likely to yield more
precise forecasts.
This algorithm operates as a classifier, employ-
ing numerous decision trees constructed from dif-
ferent subsets of the dataset. It then leverages the
mean of these predictions to improve the accuracy
of its forecasts. Unlike a single decision tree, ran-
dom forest aggregates predictions from multiple
decision trees, which significantly contributes to its
effectiveness.
However, while SVM’s application in classification discarding 4 features that were deemed less informa-
tasks is well documented, its usage in regression is tive or potentially noisy.
less prevalent in the ML literature. Nevertheless, SVR Interestingly, as we introduced additional features
emerges as a potent supervised learning algorithm beyond this selected subset, we observed a decline
designed to predict continuous values. in precision. This decline can be attributed to the
SVR shares the fundamental principles of SVM but inclusion of non-useful data that does not contribute
diverges in its objective. Instead of classifying data positively to the specific combination of features and
points, SVR seeks to identify the most appropriate classifier we are using. The careful selection of fea-
line, referred to as a hyper plane that maximizes the tures is crucial in ensuring the model’s efficiency and
encompassment of data points within a predefined predictive power (Figure 46.13).
threshold. This threshold signifies the distance In this specific combination, we were able to attain
between the hyper plane and the boundary line. a commendable maximum accuracy of 91%. This
It’s worth noting that the time complexity of SVR achievement was made possible by selecting and using
escalates significantly with an increase in the number 5 key features for the model. We deliberately excluded
of samples, making it less suitable for scaling datasets any additional features beyond these 5 during the
with a large sample size, often exceeding several tens modeling process.
of thousands. In such scenarios, linear SVR or sto- What’s noteworthy is that when we introduced the
chastic gradient descent (SGD). Regression serves as feature related to RAM into the model, we observed
a swifter alternative for implementing our prediction a drop in accuracy. This decrease in accuracy sug-
model, albeit limited to considering the linear kernel. gests that the additional data represented by RAM is
A noteworthy characteristic of an SVR model is its not relevant for this specific combination of classifier
reliance on only a subset of the training data. This is and features. In fact, it appeared to introduce noise
due to the cost function disregarding samples whose or confusion, which had a detrimental effect on the
predictions are in close proximity to their target val- model’s performance.
ues. This selective utilization of data points contrib- The output of the decision tree classifier is dis-
utes significantly to the model’s efficiency. played above. The time taken to classify the model
Result of the entire algorithm that used given in is 0.03 seconds (s) to build the model and 0.2 s test
Table 46.2. it out. Twenty cases are classified correctly out of 29
In this phase of feature selection, we initially iden- accuracy rate is 75.63 %. This is the first classification
tified and removed 5–7 specific features. As a result where all the functions are used accordingly.
of this feature elimination, our model reached a peak
accuracy of 92%, representing a significant improve-
ment in predictive performance. However, an inter- Table 46.3 Bottom seven attributes with accuracy
esting observation emerged when we introduced the
feature related to random access memory (RAM) into # of attributes Accuracy % Removal of attributes
the model. Surprisingly, the inclusion of RAM led to
10 71.428 No
a decline in accuracy. It appears that, in this particu-
lar combination of features, RAM does not contrib- 9 71.42 Battery
ute meaningfully to the predictive capabilities of our 8 71.42 Weight
model and may even introduce noise or confusion, 7 71.42 Card slot
causing a reduction in accuracy (Table 46.3). 6 75 Display
In this particular combination, we managed to 5 75 Thickness
achieve the highest level of accuracy, reaching an
4 57.14 RAM
impressive 94%. Our feature selection process led
us to retain a subset of 5–6 relevant features while
1 25 Display size
2 46.4 Memory
3 46.4 Card slot
4 75 Camera
5 75 Video
Figure 46.13 Attributes vs. accuracy of five attributes
360 Forecasting mobile prices: Harnessing the power of machine learning algorithms
This particular combination yields a maximum The trade-off between maximum accuracy and the
accuracy of 96% and all of the features chosen, with minimum number of features is a common challenge
the exception of the functions mentioned above, are in ML. Finding the right balance is often dependent
depressed. The introduction of this feature led to a on the specific problem and dataset. Ideally, a model
decrease in accuracy because it contains redundant or should use the minimum number of features neces-
irrelevant data for the specific classifier and feature sary to achieve a satisfactory level of accuracy, as this
selection algorithm. can lead to more efficient and practical solutions.
The number of variables drawn at random for each
split, the splitting rule, the minimum number of sam- Conclusion
ples a node must contain, the number of trees, and the
number of observations drawn at random for each In our research, we’ve explored various feature selec-
tree are just a few of the hyper parameters that need tion algorithms and classifiers, including combina-
to be set by the user when using the random forest tions like KNN, tree classifiers, decision trees, and
(RF) algorithm (Wenck et al., 2023). more. What’s interesting is that we’ve achieved com-
The utilization of surrogate variables holds great parable results with both feature selection and classi-
potential for effective variable selection and for fier combinations. By carefully selecting only the most
examining the intricate relationship between pre- relevant features while minimizing redundancy, we’ve
dictor and outcome variables in high-dimensional managed to attain maximum accuracy in our predic-
omics datasets (Arora, Srivastava, and Garg 2020) tions. It’s worth noting that during feature selection,
(Figure 46.14). the presence of irrelevant or redundant features in the
dataset can significantly compromise the performance
Comparative study of both classifiers in our prediction model. Conversely,
removing essential features from the dataset, espe-
In ML, model evaluations often revolve around two cially through reverse selection, can lead to decreased
primary criteria and they are as follows: efficiency. One of the primary reasons for reduced
Maximum accuracy: Achieving the highest accu- accuracy in our model is the limited number of
racy possible is a fundamental goal in machine learn- occurrences or instances in the dataset. It’s crucial to
ing. High accuracy indicates that the model is making acknowledge this limitation and its potential impact
correct predictions with a low rate of error. It implies on the model’s performance. Additionally, it’s impor-
that the model is effectively capturing the underly- tant to consider that transitioning from a regression
ing patterns and relationships in the data, leading to task to a classification task, or vice versa, can intro-
reliable predictions. However, maximizing accuracy duce more errors into the model. This highlights the
typically requires the use of more precise and relevant importance of choosing the most suitable modeling
data in the prediction model. approach for the specific problem at hand. Ultimately,
Minimum number of features: Reducing the num- our model has successfully predicted mobile prices
ber of features used in a model can have several accurately, leveraging the expressive power of their
advantages. It can lead to a more efficient and stream- features. The data gathered from the Internet for price
lined model with reduced memory requirements. prediction has proven to be accurate, contributing to
Additionally, fewer features can result in lower com- the overall reliability of our results.
putational complexity, making the model faster to
train and use in real-time applications. However, it’s
Outcome of the work
crucial to strike a balance because selecting too few
features may lead to a loss of important information Cost forecasting is a crucial aspect of both the mar-
and reduced model performance. ket and various business operations. The process of
Applied Data Science and Smart Systems 361
estimating the cost of mobile phones can be applied spectrum of scenarios and market fluctuations,
to a wide range of products, including cars, food improving the model’s ability to generalize. Ad-
items, medicines, laptops, and more. This methodol- ditionally, feature selection plays a crucial role
ogy allows for a comprehensive understanding of the in enhancing accuracy. Identifying and including
pricing dynamics across various industries. relevant features that have a significant impact
An effective marketing strategy often revolves on price prediction is essential for building a
around identifying affordable products that offer highly accurate model. Overall, a large and com-
optimal specifications. This involves finding products prehensive dataset, combined with thoughtful
with the lowest possible price (p-minimum) while feature selection, is key to achieving higher ac-
providing top-notch features. By comparing products curacy in the selected prediction model.
based on their specifications, cost, manufacturer, and
other factors, consumers can make informed purchas- References
ing decisions.
Recommendations for products within an eco- Pudaruth, Sameerchand. (2014). Predicting the price of
used cars using machine learning techniques. Int. J. Inf.
nomic range can be particularly valuable to custom-
Comput. Technol. 4(7): 753–764.
ers. This suggests products that meet their budgetary Visit, L. (2004). House price prediction: Hedonic price
constraints while still offering satisfactory features model vs. artificial neural network. Am. J. Appl. Sci.,
and quality. Leveraging company models can assist 1, 193–201.
customers in finding the ideal mobile device that suits Noor, K. and Jan, S. (2017). Vehicle price prediction sys-
their specific needs and financial considerations. tem using machine learning techniques. Int. J. Comp.
In essence, customers seeking a cost-effective Appl., 167, 27–31.
smartphone with desirable features can benefit from Saini, I. S. and Kaur, N. (2023). Comparison of various
suggested models that help them identify the best regression techniques and predicting the resale price
mobile device within their price range. This approach of cars. 2023 10th Int. Conf. Comput. Sustain. Glob.
enhances consumer satisfaction and promotes Dev. (INDIACom), 857–861.
McMenamin, J. Stuart, and Frank A. Monforte. (2000).
informed decision-making in the marketplace.
Statistical approaches to electricity price forecast-
ing. Pricing in Competitive Electricity Markets:
Future work extension 249–263.
Mishra, S., Gupta, P., Pandey, S., and Shukla, J. P. (2014).
i. A high-quality dataset is essential for building An efficient approach of artificial neural network in
a more accurate predictive model. Not all al- runoff forecasting. Int. J. Comp. Appl., 92, 9–15.
gorithms are equally suitable for every type of Gegic, Enis, Becir Isakovic, Dino Keco, Zerina Masetic, and
model, and the choice of algorithms should be Jasmin Kevric. (2019). Car price prediction using ma-
based on the characteristics of the data and the chine learning techniques. TEM Journal. 8(1): 113.
specific problem at hand. Ensuring the dataset Phyu, Thu Zar, and Nyein Nyein Oo. (2016). Performance
is clean, well structured, and representative of comparison of feature selection methods. In MATEC
the problem is critical for achieving accurate web of conferences, 42, 06002. EDP Sciences.
predictions. Listiani, M. (2009). Support vector regression analysis for
price prediction in a car leasing application (Doctoral
ii. Leveraging more advanced AI techniques can
dissertation, Master thesis, TU Hamburg-Harburg).
indeed enhance accuracy and enable more pre- Muhammad, A. and Z. Y. Khan. (2018). Mobile price class
cise predictions of product prices. Techniques prediction using machine learning techniques. Int. J.
like deep learning, neural networks, and natural Comp. Appl., 179, 6–11.
language processing can be applied to capture Kalaivani, K. S., N. Priyadharshini, S. Nivedhashri, and R.
complex patterns and relationships in the data, Nandhini. (2021). Predicting the price range of mo-
leading to improved price predictions. bile phones using machine learning techniques. In AIP
iii. Developing a dedicated software or mobile ap- Conference Proceedings, 2387(1). AIP Publishing,
plication for future predictions related to mobile 2021.
products can streamline the process and make it Nadeem, A. M., Singh, G., Badoni, P., Walia, R., Rahi, P.,
more accessible to users. Such applications can and Saddiqui, A. T. (2023). An efficient ADA boost
and CNN hybrid model for weed detection and re-
provide real-time pricing information, product
moval. 2023 10th Int. Conf. Comput. Sustain. Glob.
recommendations, and market insights, enhanc- Dev. (INDIACom), 244–250.
ing the user experience and decision-making. Balaji, V., Aishwarya, R., Sikhwal, Y., and Ramesh, S.
iv. To achieve maximum accuracy in predicting (2023). Used car price prediction using machine learn-
mobile phone prices, it’s important to include ing. Adv. Sci. Technol., 124, 512–517.
a diverse range of cases in the dataset. Adding Singh, Deepesh. (2012). The High-Quality Low-Price Busi-
more data instances can help capture a broader ness Strategy of Samsung Mobile in Penetrating Com-
362 Forecasting mobile prices: Harnessing the power of machine learning algorithms
petitive Market of India. Available at SSRN 2198366, Seifert, S., Gundlach, S. and Szymczak, S. (2019). Surrogate
1–20, [Link] minimal depth as an importance measure for variables
Singh, J., Singh, S., Singh, S., and Singh, H. (2019). Evaluat- in random forests. Bioinformat., 35, 3663–3671.
ing the performance of map matching algorithms for Brar, P. S., Shah, B., Singh, J., Ali, F., and Kwak, D. (2022).
navigation systems: an empirical study. Spat. Inform. Using modified technology acceptance model to evalu-
Res., 27, 63–74. ate the adoption of a proposed IoT-based indoor disas-
Sidhu, V. P. S., Goyal, K., and Goyal, R. (2017). An investi- ter management software tool by rescue workers. Sen-
gation of corrosion resistance of HVOF coated ASME sors, 22(5), 1866. [Link]
SA213 T91 boiler steel in an actual boiler environ- Wenck, Soeren, Thorsten Mix, Markus Fischer, Thomas
ment. Anti-Corr. Methods Mat., 64, 499–507. Hackl, and Stephan Seifert. (2023). Opening the Ran-
Segal, Mark R. (2004). Machine learning benchmarks and dom Forest Black Box of 1H NMR Metabolomics
random forest regression. 1–14. Data by the Exploitation of Surrogate Variables. Me-
Probst, Philipp, Marvin N. Wright, and Anne-Laure Boul- tabolites. 13(10): 1075.
esteix. (2019). Hyperparameters and tuning strate- Arora, Pritish, Sudhanshu Srivastava and Bindu Garg.
gies for random forest. Wiley Interdisciplinary Re- (2020). Mobile Price Prediction using WEKA. Interna-
views: data mining and knowledge discovery. 9(3): tional Journal of Science & Engineering Development
e1301. Research ([Link]), 5, 330–333.
47 Deep learning-based chronic kidney disease (CKD)
prediction
J. Angel Ida Chellama, M. Preethi, R. Rajalakshmi and E. Bharathraj
Sri Ramakrishna Engineering College, Coimbatore, Tamil Nadu, India
Abstract
In the area of healthcare applications such as classification, illness prediction, etc., there is a growing emphasis placed on the
categorization of medical data. In addition to learning, neural systems also have other advantageous traits including poor
or absent data management, such as the capacity to separate noise, vulnerability, or imprecision. Hence the significance of
feature selection is that it decreases the classifier capacity to the measurements that are considered generally pertinent in
precise classification. The primary objective of this work is to arrange the medical data and investigate the viability of using
distinctive input features and classifiers to find the medical datasets. This work proposed deep neural network (DNN) for
classification. From the result outcome, it is observed that the proposed DNN classifier produces higher accuracy, sensitiv-
ity, and specificity rates than machine learning (ML)-based classification algorithms with respect to chronic kidney disease
(CKD) dataset.
Keywords: Deep learning, chronic kidney disease (CKD), feature extraction, classification
a
[Link]@[Link]
364 Deep learning-based chronic kidney disease (CKD) prediction
Figure 47.2 Block diagram of the proposed approach for CKD prediction
Input
Result and discussion
^F: {f (1), . . . f (m)} represents ‘m’ instances of chronic
disease dataset, and xi ∈ {0,1} denotes the class Performance measures
label. Several performance measures are used to analyze the
efficiency of the proposed and existing algorithms.
Output The accuracy, sensitivity, and specificity have been
Obtain the cost function θ by considering the used to evaluate the performance of the proposed
updated momentum value and learning rate algorithm.
1. Let ε be the learning rate and φ be the
parameter of momentum
2. Let θ be the initial cost function and ve denote Accuracy
the prior velocity It calculates the global prediction rate is determined
3. Let T denote the threshold value by dividing the number of successfully classified CKD
4. while (min_stop<T) by the total number of CKD affected patients used for
5. Take into account the “m” instances from the classification, yielding the following ratio,
chronic disease dataset {f(1), . . . f(m)} with target
x(i). TP + TN
Accuracy =
6. Estimate the gradients as g ← 1 ∆θ ∑ L(f(f(i); TP+TN+FP+FN
mb
θ), x(i))
CKD dataset
TN
Sensitivity =
TP+FN
Specificity
Specificity is a metric that establishes whether or
not a person has the CKD (Table 47.1). The ratio of
correctly categorized normal person to total CKD
affected patients is what determines follows,
Figure 47.6 Comparative analysis
Sensitivity = TN
TP+FN
Table 47.3 and Figure 47.6 represent the perfor-
DNN achieved the highest performance. The accu- mance analysis of classification algorithms. From the
racy, sensitivity and specificity measure of DNN experimental results, it is observed that the proposed
were 98.89%, 98.56% and 93.23%. Table 47.2 and DNN classifier produces higher accuracy, sensitivity,
Figure 47.4 illustrate the performance analysis of the and specificity rates than other ML-based classifica-
proposed DNN for CKD dataset. tion algorithms with respect to CKD dataset.
The presence or absence of CKD is often indicated
by the two classes, class 1 and class 2, in the CKD
Conclusion
dataset. In Figure 47.5, y-axis displays performance
metrics including accuracy, sensitivity, and specific- The evaluation of the optimal subset of features
ity while the x-axis displays the class. The suggested among the multiple variables contained in the taken-
approach has a class 1 accuracy of 96.5%, a sensi- into-account CKD dataset is the crucial problem in
tivity of 95%, and a specificity of 94%. The current medical data classification. The missing values were
study also achieved 97.86% accuracy, 84.5% sensi- first eliminated during the pre-processing step of
tivity, and 97.89% specificity for class 2. the data. Afterwards, the suggested OLLP algorithm
Applied Data Science and Smart Systems 369
chose the best characteristics. Using the best subset of Rubini, L. and Eswaran, P. (2015). Generating comparative
characteristics, the dataset was then separated into 2 analysis of early stage prediction of chronic kidney
classes, which depending on the presence of CKD and disease. Int. J. Modern Engg. Res., 5(7), 49–55.
lack of CKD, [Link]. DNN algorithm was Rubini, L. Jerlin, and Perumal Eswaran. (2015). Generat-
ing comparative analysis of early stage prediction
used for this classification since it is the best suitable
of Chronic Kidney Disease. International Journal of
technique for data classification. From the result out-
Modern Engineering Research (IJMER). 5(7): 49–55.
come, it is observed that the proposed DNN classifier Singh, N. and Singh, P. (2020). A stacked generalization
produces higher accuracy, sensitivity, and specificity approach for diagnosis and prediction of type 2 di-
rates than ML-based classification algorithms with abetes mellitus. Adv. Intel. Sys. Comput., 990. doi:
respect to CKD dataset. In future, this research work 10.1007/978-981-13-8676-3_47.
will be integrated ML-based health monitoring sys- Alloghani, M., Al-Jumeily, D., Hussain, A., Liatsis, P., and
tems in the cloud platform, which have the required Aljaaf, A. J. (2020). Performance-based prediction
provision to access the disease data at any time and of chronic kidney disease using machine learning for
location. high-risk cardiovascular disease patients. Stud. Com-
put. Intel., 855. doi: 10.1007/978-3-030-28553-1_9.
Harimoorthy, K. and Thangavelu, M. (2021). Multi-disease
References prediction model using improved SVM-radial bias
Feng, L., Wang, J., Tang, B., and Tian, D. (2014). Life grade technique in healthcare monitoring system. J. Amb.
recognition method based on supervised uncorre- Intel. Human. Comput., 12(3). doi: 10.1007/s12652-
lated orthogonal locality preserving projection and 019-01652-0.
K-nearest neighbor classifier. Neurocomputing, 138, Thakur, D., Singh, J., Dhiman, G., Shabaz, M., and Gera, T.
271–282. (2021). Identifying major research areas and minor re-
Bala, S. and Kumar, K. (2014). A literature review on kid- search themes of android malware analysis and detec-
ney disease prediction using data mining classifica- tion field using LSA. Complexity, 2021, 1–28. https://
tion technique. Int. J. Comp. Sci. Mob. Comput., 3(7), [Link]/10.1155/2021/4551067.
960–967. Khamparia, Aditya, Gurinder Saini, Babita Pandey, Shrasti
Murtagh, F. E., Addington-Hall, J. M., Edmonds, P. M., Tiwari, Deepak Gupta, and Ashish Khanna. (2020).
Donohoe, P., Carey, I., Jenkins, K., and Higginson, KDSAE: Chronic kidney disease classification with
I. J. (2007). Symptoms in advanced renal disease: A multimedia data learning using deep stacked autoen-
cross-sectional survey of symptom prevalence in stage coder network. Multimedia Tools and Applications.
5 chronic kidney disease managed without dialysis. J. 79: 35425–35440.
Pall. Med., 10(6), 1266–1276. Zebari, Rizgar, Adnan Abdulazeez, Diyar Zeebaree, Dilo-
Rubini, L. and Eswaran, P. (2015). Generating comparative van Zebari, and Jwan Saeed. (2020). A comprehensive
analysis of early stage prediction of chronic kidney review of dimensionality reduction techniques for fea-
disease. Int. J. Modern Engg. Res., 5(7), 49–55. ture selection and feature extraction. Journal of Ap-
Kunwar, V., Chandel, K., Sabitha, A. S., and Bansal, A. plied Science and Technology Trends. 1(2): 56–70.
(2016). Chronic kidney disease analysis using data Fatima, Meherwar, and Maruf Pasha. (2017). Survey of ma-
mining classification techniques. Proc. 2016 6th Int. chine learning algorithms for disease diagnostic. Jour-
Conf. Cloud Sys. Big Data Engg. (Confluence), 1–6. nal of Intelligent Learning Systems and Applications.
Kunwar, V., Chandel, K., Sabitha, A. S., and Bansal, A. 9(1): 1–16.
(2016). Chronic kidney disease analysis using data Mushtaq, Zaigham, Muhammad Farhan Ramzan, Sikandar
mining classification techniques. Proc. 2016 6th Int. Ali, Samad Baseer, Ali Samad, and Mujtaba Husnain.
Conf. Cloud Sys. Big Data Engg. (Confluence), 79– (2022). Voting classification-based diabetes mellitus
110. prediction using hypertuned machine-learning tech-
Chetty, N., Vaisla, K. S., and Sudarsan, S. D. (2015). Role niques. Mobile Information Systems. 2022: 1–16.
of attributes selection in classification of chronic kid- Sornam, M. and Prabhakaran, M. (2018). Logit-based arti-
ney disease patients. Proc. 2015 Int. Conf. Comput. ficial bee colony optimization (LB-ABC) approach for
Comm. Sec. (ICCCS), 1–7. dental caries classification using a back propagation
Chen, H., Chang, P., Hu, Z., Fu, H., and Yan, L. (2019). A neural network. Integr. Intel. Comput. Comm. Sec.
spark-based ant lion algorithm for parameters optimi- Stud. Comput. Intel., 1(1), 79–91.
zation of random forest in credit classification. 2019 Singh, J. and Singh, K. (2009). Statistically analyzing the
IEEE 3rd Inform. Technol. Netw. Elec. Autom. Con. impact of automated ETL testing on the data quality
Conf. (ITNEC), 978(1), 992–996. of a data warehouse. Int. J. Comp. Elec. Engg., 1(4),
Luck, M., Bertho, G., Bateson, M., Karras, A., Yartseva, A., 488.
Thervet, E., Damon, C., and Pallet, N. (2016). Rule- Zheng, B., Zhang, J., Yoon, S. W., Lam, S. S., Khasawneh,
mining for the early prediction of chronic kidney dis- M., and Poranki, S. (2015). Predictive modeling of
ease based on metabolomics and multi-source data. hospital readmissions using metaheuristics and data
Plos One, 11(11), 1–20. mining. Exp. Sys. Appl., 42(20), 7110–7120.
48 Cattle identification using muzzle images
J. Anithaa, R. Avanthika, B. Kavipriya and S. Vishnupriya
Sri Ramakrishna Engineering College, Coimbatore, Tamil Nadu, India
Abstract
Nowadays livestock management is critical for a country’s economy, which includes identification of breed, total count of
cattle in a region, and identification of unique cattle. The livestock management is very important for the government, when
insurance claims are made during floods or epidemic. Hence advanced techniques are required to use biometrics like muzzle
images to uniquely identify the cattle. The cattle identification system’s aim is to identify individual cattle with its unique
muzzle print. Similar to finger print of human being, every individual cattle possess unique muzzle patterns. With the feature
extraction techniques, the unique extracted features could be matched against the template of cattle to identify it. The feature
matching based on YOLO V5 algorithm has obtained an accuracy of 80%, whereas the system that uses feature extraction
and deep learning methodology like SIFT, CNN and VGG16 trained with more than 200 cattle images has obtained accuracy
of 95%. The extracted features are stored and could be used for matching and identifying the cattle in future. This would
prevent many issues like false insurance claim, help the abattoir to track their cattle, etc.
Keywords: Feature extraction, muzzle pattern, VGG16, muzzle images, CNN, YoloV5 model
anitha.j@[Link]
a
Applied Data Science and Smart Systems 371
Holstein Friesian cattle and identification in agricul- algorithm for feature detection, and achieves a 96%
tural settings. It introduces new datasets and shows classification accuracy when classifying cattle muzzles
that deep learning can achieve 99.3% accuracy in into ten groups, outperforming traditional methods
cattle detection and 86.1% accuracy in individual with 90% accuracy.
identification using in-barn imagery, and 98.1% Mahmoud et al. (2021) presented a methodical
accuracy with UAV footage. These results suggest that review on deep learning applications in precision
marker-less cattle identification is feasible and robust cattle farming, emphasizing health and identifica-
in uncluttered environments, complementing existing tion. Among 678 studies, 56 meet criteria, with cattle
tagging methods. identification (58%) and health monitoring as major
Awad et al. (2019) investigated cattle identifica- applications. Convolutional neural networks (CNNs),
tion and traceability through the use of muzzle print particularly ResNet, are popular models. Challenges
photos and the bag-of-visual-words (BoVW) method. include image quality and data processing.
For feature extraction, it employs two feature detec- Qiao et al. (2021) developed a novel deep learning
tors, accelerated strong features and stable extremal approach for identifying cattle using video analysis.
regions in the BoVW model. The results demonstrate It combines Inception-V3 CNN, BiLSTM, and self-
the practicality of BoVW, with SURF achieving higher attention mechanisms to achieve 93.3% accuracy in
accuracy (up to 93%) than MSER (up to 67%) for identifying cattle from rear-view videos, surpassing
various training dataset sizes. This technology has existing methods. Additive attention outperforms
the potential to be used for cattle identification and multiplicative attention, and longer video sequences
traceability. enhance identification accuracy, offering potential for
The study by Bello et al. (2020) introduced the use automated cattle identification in precision livestock
of stacked denoising auto encoders and deep belief farming.
networks for cow nose image texture feature extrac- Shen et al. (2019) introduced a contactless cow
tion. These methods aid in animal biometrics, partic- identification approach using CNNs. It gathers
ularly cow recognition, a vital aspect of automated side-view images of cows, uses YOLO to recognize
animal registration. Experimental results indicate that objects, and fine-tunes a CNN model for individ-
the deep belief network achieves an impressive accu- ual cow classification Gera et al. (2021). With 105
racy of around 98.99% using a dataset of 4000 muz- images, the method achieves a 96.65% accuracy,
zle images from 400 cows, contributing to the field of surpassing previous experiments, indicating its effec-
animal biometrics. tiveness for cow identification and broader livestock
The research by Li et al. (2022) focused on beef applications.
cattle identification through unique muzzle patterns, Zin et al. (2018) proposed a precision dairy farm-
important for traceability and disease tracking. It col- ing, a key innovation in the fourth industrial revo-
lected a high-quality dataset of 4923 muzzle images lution, leverages IoT, AI, and cloud computing to
taken from 268 US feedlot cattle and tested 59 deep enhance cow health and farm profitability. A hybrid
learning models, achieving 98.7% accuracy with 28.3 visual stochastic approach combining image tech and
ms/image processing speed. The augmentation of data stats to monitor dairy cows for cow ID, body con-
and weighted cross-entropy loss function improves dition, estrus behavior, and calving time prediction,
accuracy, demonstrating the potential of deep learn- using Markov chains for decision-making based on
ing in precision livestock management. The dataset real-world and existing data.
is available for further research in the beef cattle
industry. Objectives
The study by Li et al. (2017) proposed an approach
for automated cow identification using tail head This paper is aimed to develop a deep learning model
images and Zernike moments as shape descriptors. that could identify individual cattle with their unique
Four classifiers were tested, with quadratic discrimi- muzzle image. This makes sure that no scam could be
nant analysis (QDA) achieving the highest accuracy at done during the claim of insurance and hassle free.
99.7% and support vector machines (SVM) achieving Another objective is to deploy the model in a mobile
the highest precision at 99.6%. QDA and SVM were application.
the most effective methods for precision animal man-
agement in dairy cow identification. Mahmoud et al. Methodology
(2015) developed a muzzle classification system for
cattle based on multiclass support vector machines The proposed method was tested with machine learn-
(MSVMs) to ensure livestock management and ing (ML) algorithms YOLOv5 model and CNN. The
product safety. It employs pre-processing techniques following algorithms are used to check which model
for image enhancement, utilizes the box-counting has the better accuracy.
372 Cattle identification using muzzle images
YOLOv5 epochs in which the model has been trained. The pre-
YOLO is ultralytics version 5 was published in June diction model has been trained for 100 epoch with a
2020 and is currently the most sophisticated object. It batch size of 32 and the learning rate was set by 0.01.
is a cutting-edge object detection model that is widely
utilized in a variety of computer vision tasks such CNN model
as video and image classification, segmentation, and Convolutional neural network (CNN) model is the
object detection. The YOLOv5 model is a significant advanced approach that is used in these days to train
improvement over previous versions of YOLO, pro- an efficient deep learning model. Collect all the data-
viding better accuracy and faster processing speeds. set of images for cattle muzzle breed detection includ-
In general, the YOLOv5 models are a good choice for ing the images for validating the model. Preprocess
mobile deployment. They are relatively lightweight the data by scaling the images to the same size and
and efficient, making them suitable for running on dividing the dataset into training, validation, and
mobile devices. They are also accurate, with a mean test sets. By using TensorFlow framework is helpful
average precision (mAP) of 58.1% on the PASCAL to build a CNN architecture which includes multiple
VOC dataset. Large number of cattle muzzle images convolutional layers to extract features from the input
with various breeds is collected and each image is images, which follows fully connected layers which
annotated by drawing the bounding boxes around the perform classification. Train the model with valida-
image. The dataset has to be annotated image which tion image that computes the loss and back propagat-
make the YOLOv5 model to extract the features from ing the error to update the network parameter, it also
the image efficiently. Tools like Robo flow could be calculates the accuracy of the model. A transfer learn-
used for annotating the image and labeling the image ing strategy with fine-tuning is used to convert VGG
in the format of “.txt”. Figure 48.1 displays the output 16 ImageNet CNN into a VGG 16 muzzle pattern
of the muzzle classification. Train the annotated data- identifier. Figure 48.2 displays the summary of the
set with the YOLOv5 model after it has been created CNN model using the pre-trained VGG16 network
which allows you to configure several hyper param- The network is developed with the flattened layer and
eters such as learning rate, batch size, and number of a dense layer, this technique will automatically extract
the features from the muzzle image. In order to do same step: acquiring the muzzle print of the cattle.
transfer learning with fine-tuning, the original VGG This can be done in two ways. You can directly take
16 ImageNet model’s final pooling and fully con- a clear picture of the muzzle image or apply ink to
nected layer must be eliminated, then it is followed the muzzle part of the cattle and try to get its print on
by the average, pooling and dense layers are added. paper. The latter method is very efficient as it provides
The first layer is the frozen convolutional layer where a clear view of the beads and ridges. The first method,
frequent image attribute is picked up from the pre- however, may have varying quality depending on the
trained ImageNet, the other portions are known as camera used.
the unfrozen layer. Overall, four distinct model train-
ing techniques were assessed. (a) Training the model Data pre-processing
from the ground-up, no pre-initialization is done in As mentioned earlier, the quality of the images cap-
the VGG-16 architecture. (b) Transfer learning with tured by the camera may exhibit variations, mak-
the pre-initialized ImageNet weights of VGG-16 ing data pre-processing a crucial phase in our cattle
where all layers are left frozen and the SoftMax layer identification system. This phase encompasses sev-
is added additionally to train the model for identifica- eral techniques aimed at improving the quality and
tion of cattle, and (c) The last convolutional layer is utility of the images, ensuring the model’s efficacy in
fine-tuned using transfer learning with the last convo- subsequent stages. In the initial step of data prepro-
lutional layer unfrozen so that it was weights could cessing, we address the presence of unwanted noise
be modified. in the images. Noise can arise from various sources,
including camera sensors and environmental factors.
Removing noise is essential to ensure that the sub-
Implementation
sequent feature extraction process is not adversely
The system generally proposes two perspectives on affected by extraneous artifacts in the images. Image
the approach. Both phases start initially with the enhancement techniques are employed to refine and
374 Cattle identification using muzzle images
clarify the visual information within the images. This in this feature vector corresponds to specific muzzle
enhancement process plays an important role in the pattern features, such as ridge patterns and texture
validation of features extracted in later stages. One details. These features, extracted from the CNN, are
notable technique we employ is contrast limited subsequently utilized for cattle identification, com-
adaptive histogram equalization (CLAHE). CLAHE paring them against stored templates in a database to
is used for improving the image quality due to poor determine if the input image matches any previously
lighting or low contrast. It operates by locally adjust- identified cattle using similarity scoring or classifica-
ing the contrast in different regions of the image, pre- tion techniques. This process translates visual infor-
venting over-amplification of noise and preserving mation into a comprehensive feature representation,
fine details. CLAHE divides the image into smaller, enabling precise and efficient identification of indi-
overlapping blocks or tiles. Within each block, the vidual cattle.
histogram of pixel intensities is equalized, ensuring a
more balanced distribution of pixel values. Adaptive Storage and retrieval
contrast limiting prevents extreme amplification of The extracted features are stored in the database.
contrast in areas with excessive noise. The results of With the help of similarity score, we can conclude
CLAHE application include improved image contrast whether two images match or not. Similarity score
and enhanced visibility of subtle features, making it calculation is an important step in many machine
particularly well-suited for biometric systems, such as learning and computer vision applications like image
cattle muzzle pattern identification. The application recognition and object detection. The similarity score
of CLAHE to images of muzzle points is pivotal for is a quantitative measure that indicates how similar
our biometric recognition system. These images, often two images, objects, or features are to each other.
used to identify animals, can present challenges due Figure 48.3 shows the enrolment and verification of
to variations in lighting conditions. CLAHE’s ability the cattle muzzle images.
to enhance contrast without amplifying noise ensures
that the unique patterns and features of the muzzle Results and discussion
are accentuated. This enhancement significantly con-
tributes to the accuracy of our recognition system. VGG16 model
This technique counts the number of blob and ridge The results obtained demonstrate the effectiveness of
regions in muzzle images that have regions where the YOLOV5 and VGG16 models in accurately iden-
blobs and ridges meet. Following pre-processing, tifying cattle based on their muzzle patterns. By using
the muzzle image is separated into various regions these models, the accuracy of cattle identification has
of interest using the texture segmentation technique
to obtain the discriminatory characteristics. The final
stage involves creating a feature vector for each cattle
muzzle image. The quality of the muzzle images is
initially evaluated during the feature extraction pro-
cess to see if it is suitable for additional processing.
Following the CLAHE technique’s improvement in
the quality of muzzle images, a notable set of features
(texture and pixel intensity features) are extracted
and represented.
Feature extraction
In the feature extraction process, preprocessed
cattle muzzle images are fed into the CNN model,
where each layer performs mathematical operations,
including convolutions, pooling, and non-linear acti-
vations. These operations progressively extract hier-
archical features, identifying patterns and structures
within the images. Outputs from selected interme-
diate layers are then collected, representing high-
dimensional image abstractions at varying levels of
detail. These intermediate outputs are combined to
create a feature vector, a numerical representation
encapsulating essential patterns recognized by the
CNN. Typically high-dimensional, each dimension Figure 48.3 Flow diagram of proposed system
Applied Data Science and Smart Systems 375
been significantly improved, and this can help farmers commonly used metric in image classification, repre-
and researchers in monitoring the health and produc- senting the ratio of images classified correctly to the
tivity of their cattle. total cattle images in the dataset.
The YOLOV5 model, a variant of the You Only
Look Once algorithm, has proven to be effective in Mobile application
object detection tasks, including cattle identification The proposed system is integrated with a mobile
using muzzle patterns. Its ability to detect objects in application for easy access. The mobile application
real-time, while maintaining high accuracy, makes it helps farmers to register their cattle in the database,
a reliable model for identifying cattle. In our experi- and the cattle can be authenticated with the help of
ments, YOLOV5 achieved a mean average preci- the mobile application. Figure 48.6 shows the GUI for
sion (MAP) of 58.1%. The MAP metric quantifies the farmers to sign up for the application by register-
the overall precision of an object detection model ing their phone number and authentication is done
across multiple categories, making it a valuable mea- using OTP.
sure of performance in multi-class tasks like cattle Once the owner is signed in, the application asks
identification. for details of the cow, including ear tag number,
On the other hand, the VGG16 model, a CNN has breed, and muzzle images. Figure 48.7 displays the
demonstrated its effectiveness in image classification dashboard for details of the cattle. These details are
tasks. With its deep architecture, it has the capability
to learn complex features and patterns that make it
suitable for identifying cattle based on their muzzle
patterns. In our experiments, VGG16 achieved an
outstanding classification accuracy of 95% and loss
as depicted in Figures 48.4 and 48.5. Accuracy is a
stored in the database and can be accessed when try- tool. Table 48.1 shows the descriptive statistics of
ing to find a match. feedback responses. With an overall mean score of 4.3
To verify a cattle, the muzzle image is captured out of 5, as shown in Table 48.1, this product stands
with the help of the camera. Figure 48.8 displays the out in every aspect.
results of the matched cattle. The captured image is The lowest feedback was regarding capturing the
then given to the model, and the final result is dis- muzzle image using the mobile camera. The average
played. If the cattle’s matching feature is found in the answer to the question “How efficient is the process
database the match is provided, if no data is available of capturing the muzzle image with the mobile cam-
then no results is provided as shown in Figure 48.9. era?” has a rating of 3.8 on a scale of 5. It implies that
All participants were then given access to a ques- participants encountered difficulties in capturing the
tionnaire to provide insight on the mobile application. muzzle image with a mobile camera, as the pixel lev-
Forty users completed the survey after employing els of each mobile’s camera might vary. Training the
every function of the mobile application. To quan-
tify general response, all questions were asked on a
5 point Likert scale (Strongly disagree: 1, Disagree: 2,
Neutral: 3, Agree: 4, Strongly agree: 5). This question-
naire has only five short sentences to analyze the mod-
el’s and mobile application’s efficiency and usefulness.
The feedback was examined using the Microsoft excel
Questions N Mean
model with more variated pixel images could resolve Andrew, William, Colin Greatwood, and Tilo Burghardt.
the issue. Figure 48.10 depicts the feedback of the (2017). Visual localisation and individual identifica-
product in scale of 0–5. tion of holstein friesian cattle via deep learning. In
Proceedings of the IEEE international conference on
computer vision workshops, 2850–2859.
Conclusion and future scope Awad, A. I. and Hassaballah, M. (2019). Bag-of-visual-
words for cattle identification from muzzle print im-
In conclusion, cattle identification using muzzle
ages. Appl. Sci., 9(22), 4914.
images is a promising method for individual animal
Bello, R.-W., Hj Talib, A. Z., and Bin Mohamed, A. S. A.
recognition. Muzzle images have unique features (2020). Deep learning-based architectures for recogni-
that can be used to distinguish between different tion of cow using cow nose image pattern. Gazi Uni-
cattle, and the use of image recognition technology versity J. Sci., 1–1.
can automate the identification process. This method Kumar, S., Pandey, A., Sai Ram Satwik, K., Kumar, S.,
has the potential to improve management practices Singh, S. K., Singh, A. K., and Mohan, A. (2018).
in the livestock industry, including tracking animal Deep learning framework for recognition of cattle us-
health, monitoring feeding patterns, and identifying ing muzzle point image pattern. Measurement, 116,
potential breeding candidates. However, there are still 1–17.
challenges to overcome, such as variations in light- Li, G., Erickson, G. E., and Xiong, Y. (2022). Individual
beef cattle identification using muzzle images and deep
ing and camera angle, and the need for large datasets
learning techniques. Animals, 12(11), 1453.
to improve accuracy. The pre-trained model VGG-16
Li, W., Ji, Z., Wang, L., Sun, C., and Yang, X. (2017). Auto-
was employed in this experiment to identify different matic individual identification of Holstein dairy cows
cattle based on muzzle images. The model was trained using tailhead images. Comp. Elec. Agricul., 142,
using different cattle muzzle images. In comparison to 622–631.
the YOLOv5 model, the VGG16 has a 95% recogni- Mahmoud, H. A. and El Hadad, H. M. R. (2015). Auto-
tion rate and computation speed of 30.3 ms/image. matic cattle muzzle print classification system using
In terms of accuracy and processing speed, the VGG multiclass support vector machine. Int. J. Image Min.,
models performed better. A weighted cross entropy 1(1), 126.
loss technique along with data augmentation should Mahmoud, Md. S., Zahid, A., Das, A. K., Muzammil, M.,
help boost accuracy for identifying cattle with less and Usman Khan, M. (2021). A systematic literature
review on deep learning applications for precision cat-
number of muzzle images. The work highlights the
tle farming. Comp. Elec. Agricul., 187, 106313.
huge potential of utilizing deep learning algorithms to
Gera, T., Singh, J., Mehbodniya, A., Webber, J. L., Shabaz,
detect unique livestock based on images of their muz- M., and Thakur, D. (2021). Dominant feature selec-
zles and to aid in livestock management. With fur- tion and machine learning-based hybrid approach
ther development and refinement, cattle identification to analyze android ransomware. Sec. Comm. Netw.,
using muzzle images could become a valuable tool for 2021, 1–22.
livestock farmers and ranchers. This technology can Noviyanto, A. and Arymurthy, A. M. (2013). Beef cattle
also aid in identifying cattle theft and unauthorized identification based on muzzle pattern using a match-
movement of livestock, which can help reduce the ing refinement technique in the SIFT method. Comp.
incidence of such crimes. The use of muzzle images Elec. Agricul., 99, 77–84. [Link]
for cattle identification can also have positive impli- compag.2013.09.002.
Qiao, Yongliang, Cameron Clark, Sabrina Lomax, He
cations for animal welfare, as individualized track-
Kong, Daobilige Su, and Salah Sukkarieh. (2021).
ing can help monitor and address health issues at an
Automated individual cattle identification using video
early stage. However, it is important to ensure that the data: a unified deep learning architecture approach.
technology is implemented ethically and with proper Frontiers in Animal Science. 2: 759147.
safeguards to protect animal privacy and prevent mis- Shen, W., Hu, H., Dai, B., Wei, X., Sun, J., Jiang, L., and
use of the data. In summary, cattle identification using Sun, Y. (2019). Individual identification of dairy cows
muzzle images is a promising area of research that based on convolutional neural networks. Multimedia
could have far-reaching benefits for both the livestock Tools Appl., 79(21–22), 14711–14724.
industry and animal welfare. Zin, Thi Thi, Cho Nilar Phyo, Pyke Tin, Hiromitsu Hama,
and Ikuo Kobayashi. (2018). Image technology based
cow identification system using deep learning. In Pro-
References ceedings of the international multiconference of engi-
Ali Ismail, A., Zawbaa, H. M., Mahmoud, H. A., Abdel, neers and computer scientists, 1, 236–247.
H., Fayed, R. H., and Hassanien, A. E. (2013). A ro-
bust cattle identification scheme using muzzle print
images. Federat. Conf. Comp. Sci. Inform. Sys., 529–
534.
49 Simulation-based evaluating AODV routing protocol
using wireless networks
Bhupal Arya1,a, Dr. Jogendra Kumar2, Dr. Parag Jain3, Preeti Saroj4,
Mrinalinee Singh5 and Yogesh Kumar6
Department of Computer Science and Engineering, Roorkee Institute of Technology, Uttarakhand, India
1,3,4,5,6
2
Department of Computer Science and Engineering, GBPIET Ghurdauri Pauri Garhwal Uttarakhand, India
Abstract
For wireless ad hoc networks to function effectively and dependably, routing methods must be evaluated and enhanced.
The ad hoc on-demand distance vector (AODV) routing protocol is the main topic of this study because of its popularity
and adaptability to dynamic contexts. Simulation-based evaluation with performance metrics has been used to evaluate
the protocol’s performance accurately. This study develops a complete simulation framework to simulate different network
circumstances and situations. A number of performance metrics, such as total bytes sent, total packets sent, first packets
sent, last packets sent, first packets received, total bytes received, total packets received, last packets received, average jitter,
average end-to-end delay, and throughput, are used to assess the effectiveness of the AODV routing protocol. The simulation
scenarios take into account different node densities, traffic loads, and mobility patterns in order to give a comprehensive
evaluation of the behavior of the protocol.
Keywords: Wireless networks, simulation tool, AODV routing protocol, performance metrics
bhupalarya@[Link]
a
Applied Data Science and Smart Systems 379
• Temporal coordination and delivery reliabil- plays a crucial role in preventing the occurrence
ity: Temporal coordination and delivery reliability of loops.
have been examined by Lee and Gerla (2001), who • Route reply generation: Once RREQ packet
investigate AODV’s temporal benchmarks, includ- reached either the final destination node or either
ing “First Packet Sent Time” and “First Packet a node that possesses a valid route to the final
Received Time.” Their research underscores the destination, a frequent route reply packet (RREP)
protocol’s efficiency in promptly initiating and con- is formulated. The RREP is then unicast through
cluding transmissions. On the other hand, (Brar et the reverse path set, a compilation of nodes that
al., 2022; Casetti et al., 2002) delve into AODV’s collectively remember the path back to the source.
“Average End-to-End Delay” metric, elucidating • Maintenance of routes: Nodes sustain a continu-
its implications for reliable data delivery (Zhang et ous exchange of Hello messages to monitor the
al., 2016; Khattar et al., 2020; Garcia et al., 2022). status of links. Should a link failure or disruption
• Jitter analysis and throughput optimization: In in the route arise due to factors such as node mo-
the realm of delivery reliability and jitter analy- bility, a route error packet (RERR) is generated.
sis, Azzouni et al. (2006) assess “Average Jitter” This packet informs affected nodes to modify
as a measure of the uniformity of packet delivery their routing tables, thereby steering clear of the
timing, emphasizing its role in maintaining con- compromised route.
sistent data delivery patterns. Tang et al. (2010) • Forwarding of data: Subsequent to the establish-
extend the evaluation to throughput optimiza- ment of a route, the source node can efficiently
tion, exploring how AODV’s “Throughput” met- send data packets to the destination using the es-
ric influences its capacity to manage data traffic tablished pathway. Intermediate nodes effectively
efficiently, thus enhancing network performance forward data packets by referencing their routing
(Liu et al., 2017; Wang et al., 2018). tables.
• Holistic evaluations and multi-dimensional in- • Route expiration: Routes have a pre-determined
sights: Holistic evaluations that encompass a lifespan. If a route remains inactive for a desig-
range of metrics have been conducted by Maltz nated period, it is deemed obsolete and subse-
and Broch (1999), demonstrating the interplay quently discarded.
between “Total Bytes Received,” “Total Packets • Through these mechanisms, the AODV protocol
Received,” and “Average Jitter.” Through such tackles the intricacies of routing within dynamic
multi-dimensional analyses, they reveal AODV’s and resource-constrained network environments,
ability to achieve effective data acquisition while showcasing its ability to provide efficient and
mitigating delivery timing variations (Gupta et responsive data communication (Johnson et al.,
al., 2018; Wang et al., 2018). 2016; Kim et al., 2017; Martinez et al., 2018; Wu
et al., 2020; Garcia et al., 2022).
Ad hoc on-demand (AODV) distance vector
routing protocol Simulation setup and performance metrics
• AODV routing protocol has gained significant Table 49.1 Parameters list (simulation setup).
popularity within WSN and mobile ad hoc net-
Parameters Values
works (MANETs) due to its effective manage-
ment of network resources. AODV operates un- Area 1500 m × 1500 m
der a reactive paradigm, dynamically establishing
Channel frequency 2.4 GHz
routes between nodes as needed. This approach
effectively minimizes the burden of control mes- Model (fading) Rayleigh
sages and optimally utilizes available resources. Battery model (Mica Linear simple model
Presented below is an outline of the operational Motes)
principles of the AODV protocol. No. of nodes 100 nodes
• Route discovery: When a node aims to dispatch Node placement model Random waypoint model
data to a destination node but lacks route infor- Routing protocols AODV
mation, it triggers the method of discovery route Shadowing model Constant energy model
process.
Simulation time 900 seconds
• Propagation of (RREQ) route request: The source
node disseminates a route request packet contain- Terrain file Digital elevation model
(DEM)
ing essential particulars, including the source and
final destination addresses, along with a distinc- Traffic source Constant bit rate (CBR)
traffic load
tive sequence number. This sequence number
380 Simulation-based evaluating AODV routing protocol using wireless networks
Metrics Equations
Abstract
India has become the most populous country in the world and food is an essential necessity for human beings. The major
source of food production is agriculture. In addition, agriculture is India’s largest sector of employment. But widely tradi-
tional methods are used in agriculture that does not provide a great deal of efficiency. A solution system is deployed by the
use of machine learning (ML) which will contribute to the improvement of the agricultural sector. The proposed solution
will provide the best crop for seeding using specific traits. On the basis of soil data collected, it will suggest the necessary
fertilizer which can be used in the field. The system will assist in irrigation scheduling, which will also inform when to irrigate
the field. This will result in the saving of a significant amount of groundwater and freshwater, which is currently a concern.
Using specific inputs, the model will also provide the soil moisture of the field.
Keywords: Machine learning, crop recommendation, fertilizer recommendation, irrigation scheduling, soil moisture levels
a
triptmann11@[Link]
388 Smart agriculture using machine learning algorithms
Veenadhari et al. (2014) developed a website to After collecting dataset, the process starts with
search the effect of parameters on production of crops loading the external dataset. Firstly, the target for a
in Madhya Pradesh, India. The crops selected could model will be defined and then we perform splitting
be wheat, paddy, soybean and maize. They used the of data into train and test sets. The further step is to
decision tree algorithm for methodology. apply classification algorithms and the best accuracy
Pavan Kumar (2022) conducted the study to cat- is provided by SVC which is defined as – Let us sup-
egorize the crop. They implemented random forests pose a random point A and examine whether it lies
for improving yield production. This model provided on the left side of the plane (negative) or the right
the least mean squared error and greatest R2 value one (positive). After that, make a vector x which is
among all other regression algorithms. perpendicular to hyperplane. Consider vector x from
Jhajharia and Mathur (2022) suggested the ML origin to decision boundary is at distance “c”. Then
model implementation in the agriculture field in some project A vector on x. Thus, decision rule for this will
previous years. Out of various algorithms deployed, be defined as:
neural networks and SVMs are found to provide
®®
more precision (Thakur et al., 2021). A .x – c ≥ 0
Features Description
value of the whole column. It also includes remov- predicted for the respective algorithms. The flow chart
ing the unnecessary columns such as id number. After for this whole process is given Figure 50.3.
this, Exploratory data analysis (EDA) takes place
using various visualization libraries. Then, for imple- Soil moisture prediction
mentation of model, firstly encoding of these categori- The process starts with collecting the data and load-
cal values is done by using sklearn library. The further ing the respective dataset which consists of several
step is to split 80% data into training dataset and features which are summarized in Table 50.4.
20% into testing dataset and then training the model Here the amount of soil moisture will be the target
using various classification algorithms of machine variable for the model. Then, the process continues
learning which involves logistic regression, Gaussian with performing data wrangling and exploratory
Naïve Bayes classifier, SVC. Thus, the output will be data analysis. The further step is to implement models
390 Smart agriculture using machine learning algorithms
Table 50.3 Features description for dataset of irrigation scheduling.
Features Description
Table 50.4 Feature description for dataset of soil moisture Some famous regressors like XG boost regressor are
prediction. employed which helps in predicting the soil mois-
ture levels. The flow chart for the same is given in
Features Description
Figure 50.4.
Time Specific time at which data is
collected Results
sm Soil moisture
The solution developed helps in making agriculture
pm1, pm2, pm3 Particulate matter
more sustainable by recommending the best suitable
temp Temperature of the area in degree crop required to be sown, telling us the necessary fer-
celsius
tilizer required in the field, predicting the moisture
humd Humidity of the area where field present in soil using specific traits and help in irri-
is situated
gation scheduling. In crop recommendation system,
pres Pressure in atmosphere multiple algorithms were used. For instance, logis-
tic regression gave the accuracy of 96.36%, how-
ever, SVC algorithm provided the best accuracy i.e.
and training it by using various regression algorithms 98.63% and important classification metrics for it are
as the amount of soil moisture will be continuous. given in Table 50.5.
Applied Data Science and Smart Systems 391
Table 50.5 Classification metrics for crop Table 50.6 Classification report for fertilizer
recommendation system. recommendation system.
In fertilizer recommendation system, the algorithms R2-score was calculated for all regressors in which
logistic regression and random forest both provided score given by decision tree regressor was 95.83%.
90% accuracy in predicting the results for fertilizer. However, XG boost algorithm gives slightly different
Hence, the required fertilizer is predicted using deci- R2-score of 95.84% which is considered as the best
sion tree that provides best results with accuracy of among others and important regression metrics for it
95% and its classification report is displayed in Table are given in Table 50.7.
50.6. In irrigation scheduling, various classification
In soil moisture prediction, various regression algo- algorithms are used like Gaussian Naïve Bayes,
rithms have been used which includes linear regres- SVM, etc., out of which Gaussian Naïve Bayes
sor, XG boost regressor, decision tree regressor. The provided the accuracy of 86.99% whereas SVM
392 Smart agriculture using machine learning algorithms
gave best accuracy of 90.93% and its classification Zhou, Yan, and Murat Kantarcioglu. (2020). On transpar-
report is in Table 50.8. ency of machine learning models: A position paper. In
AI for Social Good Workshop. 1–5.
Nischitha, K., Vishkarma, D., Mahendra, N., Ashwini, and
Conclusion and future scope Manjuraju, M. R. (2020). Crop prediction using ma-
Conclusion chine learning approaches. Int. J. Engg. Res. Technol.,
23–26.
This paper proposes a system which comprises of set
Bondre, D. A. and Mahagaonkar, S. (2019). Prediction of
of solutions that are developed using ML models.
crop yield and fertilizer recommendation using ma-
This will help in increasing the crop yield and in sav- chine learning algorithms. Int. J. Engg. Appl. Sci. Tech-
ing the environment. The farmers will get to know nol., 4(5), 371–376.
the best crop they should sow for having maximum Thakur, D., Singh, J., Dhiman, G., Shabaz, M., and Gera,
profit. This system will intensely lower the farmers T. (2021). Identifying major research areas and mi-
costs by only telling them what fertilizer they should nor research themes of android malware analysis
use. It will also assist in minimizing use of excess fer- and detection field using LSA. Complexity, 2021,
tilizers which is very harmful for our environment. It 1–28.
will save freshwater which is only 3% of water pres- Prakash, S., Sharma, A., and Sahu, S. S. (2018). Soil moisture
ent on earth by soil moisture detection and irrigation prediction using machine learning. 2018 Second Int.
Conf. Invent. Comm. Comput. Technol. (ICICCT),
scheduling.
1-6. IEEE, 2018.
Ritesh, D., Dash, D. K., and Biswal, G. C. (2021). Classifi-
Future scope cation of crop based on macronutrients and weath-
er data using machine learning techniques. Results
The considered datasets have been previously col- Engg., 9, 100203.
lected by trustable sources. The system will become Kasara, Y. M. R., Kovvada Rajeev, L. N., and Sai Nandan,
more precise by adding new and extensive data from N. (2020). IoT based smart agriculture using machine
several GPS spots. More models can be developed in learning. 2020 Sec. Int. Conf. Invent. Res. Comput.
future like disease prediction. The system can further Appl. (ICIRCA), 130–134. IEEE, 2020.
extended as per user and administrative requirements Veenadhari, S., Misra, B., and Singh, C. D. (2014). Machine
to encompass other aspects of the model. learning approach for forecasting crop yield based on
climatic parameters. 2014 Int. Conf. Comp. Comm.
Informat., 1–5. IEEE 2014.
Acknowledgments Pavan Kumar, D. (2022). Smart farming through machine
We extend our sincere appreciation to all individuals learning - A review. Int. J. Res. Publ. Rev., 3(11), 792–
797.
and organizations which assist us a lot to complete
Jhajharia, K. and Mathur, P. (2022). A comprehensive re-
this research project. The assistance, uplifting and
view on machine learning in agriculture domain. IAES
support of these people is really irreplaceable. Int. J. Artif. Intel., 11(2), 753.
References
FAO, IFAD, UNICEF, WFP, WHO. (2018). The state of food
security and nutrition in the world. Building resilience
for peace and food security. Rome, FAO (2017).
51 Cloud computing empowering e-commerce innovation
Zinatullah Akramia and Gurjit Singh Bhathal
Department of Computer Science and Engineering, Punjabi University, Patiala, India
Abstract
As of the recent technology, cloud computing has become as a significant driver of innovation, particularly on the e-com-
merce industry. Its influence on this sector has been profound. This research paper delves into the transformative impact
of cloud computing on e-commerce innovation. It sheds light on how cloud computing empowers businesses to overcome
obstacles, harness advanced technologies, improve customer experiences, and stimulate growth. Through a comprehensive
analysis of the strategic use of cloud computing in fostering innovation within the e-commerce landscape, the current study
unveil the pivotal role it plays in enabling online businesses to embrace some fresh approaches, adapt to ever-changing mar-
ket, and thrive in the digital era.
Keywords: Cloud computing, customer experience, digital transformation, e-commerce, technology adoption
a
zinatullahakrami@[Link]
394 Cloud computing empowering e-commerce innovation
network of edge locations to efficiently deliver web encompasses a wide range of business transac-
content. This service automatically directs requests tions, administrative operations, and information
for web content in the Amazon Cloud to the nearest exchanges that are facilitated through various infor-
edge location. However, it is worth mentioning that mation and communications technologies. Businesses
e-commerce websites often require frequent access to can leverage a scalable and flexible infrastructure that
large-scale backend databases, and the CloudFront seamlessly supports e-commerce activities. The con-
service does not directly aid in database access (Wang vergence of cloud computing and e-commerce revo-
and Jian, 2016). lutionizes the business landscape, enabling businesses
Another solution that enhances cloud application to leverage advanced cloud-based resources and fos-
performance is the Microsoft SQL Azure Data Sync ter innovation. The integration of cloud computing
Service. With the Microsoft SQL Azure approach, into e-commerce operations has unleashed a wave of
developers can geographically distribute data to one innovation, revolutionizing traditional business mod-
or more SQL Azure data centers worldwide, utiliz- els and empowering businesses to optimize opera-
ing the data sync service. This approach allows for tional efficiency, scale seamlessly, and meet evolving
effective data distribution and synchronization across customer demands (Hao et al., 2013).
multiple locations (Wang and Jian, 2016). This seamless integration has propelled advance-
Due to some cloud computing disrupts the con- ments across the four broad categories of e-commerce
ventional network architecture model by providing transactions: business-to-business (B2B), business-to-
the flexibility and cost-effectiveness. It eliminates the consumer (B2C), consumer-to-consumer (C2C), and
negative consequences and impact of single computer consumer-to-business (C2B). Each category repre-
equipment failures, safeguarding and ensuring users sents a distinct market segment with its own dynam-
are unaffected by issues such as inaccessible devices ics and specific requirements, and cloud computing
or data loss, resulting in greater reliability and acces- has emerged as a transformative force, driving agility,
sibility to resources. Through the utilization of cloud cost-effectiveness, and technological advancements
computing, users can overcome these limitations and across the online marketplace.
experience enhanced reliability and accessibility to In today’s digital age, it is commonplace for major
their resources (Rao et al., 2013). retail companies to establish a robust online pres-
ence through websites and e-commerce platforms.
E-commerce This paradigm shift has unlocked the vast potential of
The process of purchasing essential commodities e-commerce, empowering businesses to expand their
that we often require be time-consuming, costly, and reach, streamline transactions, and generate new rev-
involve unnecessary expenditures. However, the tech- enue streams. Cloud computing serves as the corner-
nology has helped the new fortunes by the shape of stone of these online operations, providing the critical
electronic commerce, commonly known as e-com- infrastructure, storage, and computing resources nec-
merce. E-commerce helps individuals the opportunity essary for seamless e-commerce experiences (Faccia et
of buying and selling of goods and services over the al., 2016).
internet. It provides customers, partners, and other E-commerce systems help retailers by provid-
individuals to engage in a wide range of transactions ing both commercial information (such as the price
capabilities and access different services (Krypa and of products, availability of quantities, and product
Anni, 2016). reviews) and facilitating various commercial actions
The emergence of e-commerce has significantly (such as buying, selling, and returning products). As
reduced the costs associated with some different technology is growing rapidly, the exponential growth
enterprise product development and production. and use of information technology in this area has led
Moreover, it has also led to a substantial decrease to fundamental changes in the way these commercial
in circulation of commodities. This shift towards activities are performed. E-commerce has become one
e-commerce has resulted in some efficiency and con- of the main business transaction methods between
venience, benefiting both businesses and the buyer’s online merchants and consumers due to the conve-
alike (Shi et al., 2017; Singh et al., 2019). nience and efficiency it offers (Baghdadi, 2013).
Cloud computing has brought some new features The popularity of mobile communication technol-
and transformation in the field of technology, which ogy and Wi-Fi has led to the rise of mobile e-com-
enables and empowers innovation in the field of merce, which allows users to perform all kinds of
e-commerce. By harnessing the power and capabilities e-commerce activities through their mobile devices. In
of cloud computing, businesses could leverage a scal- a mobile e-commerce model based on mobile cloud,
able and flexible infrastructure that seamlessly sup- e-commerce companies do not need to build their
ports e-commerce activities. Based on the definition own service platforms. Instead, they can quickly and
of the Electronic Commerce Association, e-commerce easily operate their e-commerce processes by simply
Applied Data Science and Smart Systems 395
Figure 51.1 Shows global ecommerce retail sales reached $5.7 trillion in 2022. This share is expected to increase by
10% in 2023 and reach $6.3 trillion.
leasing cloud services on demand from cloud service provisioning and release with minimal management
providers. Consumers, on the other hand, only need effort (Rao et al., 2013).
to have a simple mobile device to access the mobile Cloud computing, a transformative force shaping
cloud and enjoy the services of the e-commerce plat- the digital landscape, has become an indispensable
form (Li et al., 2019) (Figures 51.1 and 51.2). tool for businesses of all sizes seeking to maintain
their competitive edge. This convergence of distrib-
Cloud computing uted, parallel, grid, and virtualization technologies
The rapid development of the internet has led to the empowers businesses to access computing resources
growth of many trends that are based on the internet, on-demand, eliminating the need for substantial hard-
such as cloud computing. Cloud computing is a term ware and software investments, resulting in cost sav-
with multiple definitions, one of the most renowned ings and enhanced scalability (Liu, 2011).
being IBM’s description as “a pool of virtualized com- In the dynamic realm of e-commerce innova-
puter resources that can be rapidly provisioned and tion, cloud computing unveils three distinct service
released, managed through a centralized dashboard.” models: Infrastructure Cloud, Platform Cloud, and
Cloud computing revolutionizes how businesses Application Cloud, each tailored to specific needs and
access and utilize resources, providing them with the facilitating diverse transactions. Infrastructure cloud
necessary tools precisely when needed, eliminating primarily focuses on providing users with computing
the need for upfront infrastructure investments and and storage resources, complemented by authoriza-
associated costs. This transformative approach has tion services.
unleashed significant time and cost savings, enabling Its core function involves virtualizing computing
businesses to optimize operations and enhance their and storage resources in one or multiple data centers,
agility and responsiveness to market fluctuations (Liu, enabling flexible resource allocation. Notable exam-
2011). ples of this model include Amazon’s elastic compute
Cloud computing redefines resource access and uti- cloud and IBM’s Blue Cloud. These cloud comput-
lization, enabling users to tap into a shared pool of ing models play a significant role in enabling inno-
computing resources on-demand, facilitating rapid vation within the e-commerce sector. By leveraging
396 Cloud computing empowering e-commerce innovation
Infrastructure Cloud, businesses can efficiently man- This cost-effective approach maximizes return on
age and scale their computing and storage resources. investment, empowering businesses to thrive in the
Platform Cloud, on the other hand, empowers ever-evolving e-commerce landscape.
developers to create cutting-edge applications with- Accessibility and availability: Cloud-based e-com-
out worrying about the underlying infrastructure. merce platforms democratize the online marketplace,
Together, these cloud computing models contribute to empowering businesses to transcend geographical
the advancement of e-commerce innovation, foster- boundaries and provide customers with ubiquitous
ing growth and enabling novel business opportunities access to their products. This seamless and borderless
(Treesinthuros, 2012). experience fosters a sense of connection and conve-
Another essential cloud computing model rel- nience for customers, enabling them to shop and pur-
evant to e-commerce innovation is the Application chase goods effortlessly, regardless of their location or
Cloud, which directly caters to end software users, device. As a result, cloud-based e-commerce platforms
often in the form of Software as a Service (SaaS). The ignite a surge in sales growth, elevate customer satis-
Application Cloud serves as a platform where users faction, and propel businesses to new heights of suc-
can customize, configure, assemble, install, and test cess in the ever-evolving e-commerce landscape.
each module of a software system. This level of flex- Data storage and backup: Cloud storage services
ibility empowers end users to obtain software systems empower businesses to securely and efficiently safe-
that precisely meet their needs and fulfill their specific guard their e-commerce data, ensuring business
requirements. In this model, applications such as Sales continuity and safeguarding against data loss or cor-
Force CRM, Google Apps, and Zoho have emerged ruption. With robust security measures, advanced
as highly valuable tools. These applications exemplify encryption techniques, and seamless disaster recov-
the capabilities of the Application Cloud, providing ery capabilities, cloud storage provides businesses
users with versatile and adaptable software solutions with the peace of mind to focus on growth and
that can be tailored to their unique preferences and innovation.
business demands (Liu, 2011).
Leveraging the power of sentiment analysis and Indirect role of cloud computing in e-commerce
employing a fuzzy cloud-based model, the proposed Enhanced performance and scalability: Cloud com-
system provides invaluable assistance to users in the puting unleashes a surge of computational power for
complex task of product selection. It facilitates the e-commerce businesses, enabling them to seamlessly
identification of optimal products that align with navigate traffic spikes, process vast datasets with
users’ individual preferences and requirements, while ease, and deliver exceptional performance to custom-
also incorporating the collective sentiment expressed ers. This translates into lightning-fast loading times,
by fellow customers. Through the integration of enhanced user experiences, and unwavering customer
these advanced technologies, the system enhances satisfaction, propelling businesses to the vanguard of
the decision-making process, empowering users to the ever-changing e-commerce landscape.
make informed choices amidst a vast array of product Advanced analytics and personalization: Cloud com-
options available across multiple e-commerce plat- puting unlocks a limitless reservoir of computational
forms (Yang et al., 2023). power for e-commerce businesses, enabling them to
seamlessly navigate traffic spikes, effortlessly process
Direct role of cloud computing in e-commerce vast datasets, and deliver exceptional performance to
Cloud-based infrastructure: Cloud computing customers. This translates into lightning-fast loading
empowers e-commerce businesses with scalable and times, enhanced user experiences, and unwavering
flexible infrastructure for hosting their applications customer satisfaction, propelling businesses to the
and websites, eliminating the need for costly hard- vanguard of the ever-changing e-commerce landscape.
ware investments and enabling seamless resource
scaling based on demand. This robust infrastructure Collaboration and integration: Cloud computing
ensures reliable and efficient performance for e-com- empowers e-commerce businesses to connect their
merce platforms, ensuring seamless customer experi- systems with a vast network of third-party services,
ences and operational excellence. fostering effortless integration across the entire
Cost reduction: Cloud computing fosters cost-effi- ecosystem. This seamless integration streamlines
ciency by eliminating the need for upfront investments operations, elevates customer experiences, and fuels
in physical infrastructure, maintenance, and software business growth.
licensing. By embracing a pay-as-you-go model, busi- Innovation and experimentation: Cloud comput-
nesses can dynamically align their IT expenses with ing offers a platform for e-commerce businesses to
actual usage, optimize resource allocation for other experiment with new ideas, test new features, and
growth initiatives, and enhance budgeting processes. quickly deploy innovations. This fosters a culture of
Applied Data Science and Smart Systems 397
innovation and enables businesses to stay competitive eliminating the need for in-house installation and
in a rapidly evolving market. maintenance. Cloud vendors manage a vast pool of
computing resources, allocating specific resources
Cloud computing with e-commerce to each client. This distributed cost structure makes
Mastering the intricate dance of cloud computing cloud services more affordable for retailers, expand-
and e-commerce regulations requires a deep grasp of ing accessibility to a broader range of users (Xiaofeng
their intertwined legal frameworks. While both indus- et al., 2013).
tries can function autonomously, their true brilliance The foundation of e-commerce rests upon com-
emerges when they seamlessly converge. This syner- puter networks, traditionally demanding substantial
gistic fusion unleashes a symphony of innovation, investments in hardware and software. However,
transforming the digital commerce landscape. the advent of cloud-based e-commerce has revolu-
This convergence empowers businesses to leverage tionized this landscape, significantly reducing the
the scalability, security, and agility of cloud computing need for upfront hardware and software expendi-
to drive innovation and achieve success in the dynamic tures. By leveraging cloud services, businesses can
e-commerce landscape (Xuecong et al., 2021). reap the benefits of professional maintenance at
The intertwined nature of cloud computing and a lower cost or even for free. This drastic reduc-
e-commerce underscores the necessity of their inte- tion in enterprise investment costs not only benefits
grated utilization to maximize efficiency and achieve businesses financially but also promotes the overall
desired outcomes. When organizations deploy an development of e-commerce enterprises (Shi et al.,
e-commerce system within a cloud computing environ- 2017).
ment, conducting a thorough risk assessment becomes The cost-effectiveness of cloud computing in e-com-
paramount. This evaluation is critical for identify- merce plays a pivotal role in helping enterprises mini-
ing and implementing appropriate security measures mize expenses. Instead of investing in-house software
that safeguard the e-commerce system’s integrity and development, companies can utilize the vast reposi-
ensure its seamless operation (Li et al., 2019). tories of cloud service providers to access essential
Cloud computing in e-commerce refers to “the enterprise management software. This is an effectively
policy of paying for a specific bandwidth and storage caters to more customers while ensuring a relatively
space on a scale based on the usage”, as it is far away secure environment for data storage and management
different than the traditional method, where the user (Shi et al., 2017).
were paying for a certain amount of hard disk space
and bandwidth (Taherkordi et al., 2018). 2. Speed of operations
Cloud computing heralds a new era of e-commerce, Cloud platforms excel in harnessing vast computa-
empowering businesses to elevate their operations tional resources, enabling clients to experience light-
and achieve a competitive edge (Li et al., 2019). ning-fast and efficient operations. This advantage is
Cloud computing offers a significant advantage for particularly beneficial for e-commerce platforms, as
e-commerce businesses by optimizing costs. Unlike the streamlined installation and execution process
traditional brick-and-mortar stores, e-commerce busi- offered by cloud platforms significantly reduces the
nesses can leverage cloud computing’s utility-based, time and effort required to get started. The cloud
on-demand model, paying only for the resources they environment already possesses the requisite IT infra-
use. This flexibility allows e-commerce websites to structure to host the application, minimizing the
reduce costs during periods of lower traffic, enhanc- need for clients to invest in their own infrastructure.
ing their overall cost-effectiveness (Taherkordi et al., This, in turn, expedites the execution time of vari-
2018). ous application modules, leading to enhanced overall
The symbiotic relationship between cloud com- efficiency. Furthermore, vendors expertly handle the
puting and e-commerce has fostered mutual ben- setup process, allowing clients to seamlessly utilize
efits, particularly in the realm of cost reduction for the resources as per their specific requirements (Rao
storing vast amounts of business data. By leveraging et al., 2013).
cloud data centers, companies can significantly mini- Cloud computing’s robust data processing capa-
mize the expenses associated with data storage. The bilities empower businesses to seamlessly scale their
advantages of cloud computing for e-commerce plat- computing resources in real-time, effortlessly aligning
forms are numerous, and briefly highlighted a few key with fluctuating demands. This agility enables busi-
benefits. nesses to tackle previously intractable tasks, unlock-
ing new avenues for growth and innovation. Cloud
1. Cost-effective computing’s ability to optimize resource utiliza-
Cloud computing’s pay-per-use model has revolu- tion and eliminate costly overprovisioning enhances
tionized cos-effectiveness for e-commerce businesses, operational efficiency and cost-effectiveness, driving
398 Cloud computing empowering e-commerce innovation
economic benefits for organizations of all sizes (Shi challenges to the security of e-commerce systems.
et al., 2017). Ensuring the security of e-commerce systems relies
on continuous risk assessment and management
3. Scalability throughout the system’s life cycle. This involves iden-
Scalability is a key feature of cloud computing that tifying assets, identifying new threats, and mitigating
makes it so attractive to businesses. The ability to eas- relative risks to maintain security within acceptable
ily increase or decrease resources as needed empow- limits without compromising system operations.
ers businesses to optimize costs and always have the Cloud computing’s centralized management style
resources they need. This flexibility is especially valu- amplifies the potential consequences of a security
able for businesses with fluctuating demands. breach, posing a greater risk to users. The openness
Cloud computing’s scalability makes it an adapt- and complexity of cloud environments make securing
able and cost-effective solution for businesses of all e-commerce systems based on cloud computing more
sizes. This scalability feature is particularly beneficial challenging compared to traditional network environ-
for retailers, as it helps manage the expenses associ- ments (Al-Jaberi, 2015).
ated with hosting and maintaining their platforms. Cloud computing security exhibits more complex
Additionally, scalability improves the load time of manifestations due to the virtualization and service-
applications, ensuring optimal performance even dur- oriented nature of cloud computing. In cloud envi-
ing periods of high traffic. Therefore, from an eco- ronments, user data and computations are executed
nomic standpoint, retailers using cloud computing and controlled by the cloud computing center, mak-
can greatly benefit from the scalability functionality ing it difficult for users to effectively manage them.
it offers (Rao et al., 2013). Auditing user behavior becomes essential to ensure
the successful implementation of security risk pre-
4. Security vention and control measures (Li and Junfeng,
Traditional business trends have been plagued by 2020).
numerous weaknesses, particularly in terms of secu- The Gartner report highlights seven major secu-
rity. The concerns surrounding information loss and rity risks associated with current cloud comput-
network intrusion have been effectively addressed ing technologies. These risks include privileged user
with the introduction of various standards estab- access, auditability, data location, data isolation, data
lished by organizations like ISO for cloud vendors. recovery, support for surveys, and long-term survival.
Only vendors that adhere to these standards are These vulnerabilities indicate that data and services
authorized to provide cloud services. Additionally, are susceptible to attacks within a cloud computing
customers have become more knowledgeable about environment. These attacks exploit security vulner-
the concept of cloud computing, leading them to abilities stemming from the use of cloud servers and
choose certified vendors endorsed by such organi- technologies by cloud users. As a result, the security
zations. In terms of security, the backup of data is risks in cloud-based e-commerce are significantly
a noteworthy aspect of cloud computing. Data is higher compared to traditional e-commerce.
stored in multiple locations, ensuring its protection The security system of e-commerce based on cloud
and reducing the risk of loss. Furthermore, cloud computing comprises several layers, including a net-
computing data centers are strategically located in work service layer, an encryption technology layer,
different geographical areas, adding an extra layer of a security authentication layer, a transaction proto-
security. Consequently, we can confidently state that col layer, and a business system layer. If any of these
our data with cloud providers is highly secure, allevi- security requirements are compromised, the entire
ating any concerns about potential data loss (Li and system becomes vulnerable to risks (Li and Junfeng,
Junfeng, 2020). 2020).
Cloud-based e-commerce offers enterprises a
dependable and secure data storage center, enhancing 6. Environmental impact reduction
management efficiency with a professional, safe, and By incorporating sustainable manufacturing prac-
reliable approach. By renting cloud computing serv- tices, e-commerce SMEs can minimize their carbon
ers, enterprises gain access to highly stable and reli- footprint, reduce energy consumption, and promote
able services, ensuring uninterrupted operations and responsible sourcing. The integration of cloud com-
optimal performance (Almarabeh et al., 2019). puting empowers small and medium-sized enterprises
(SMEs) to monitor and analyze their environmental
5. Risk assessment impact in real-time. This real-time visibility facilitates
While cloud computing presents significant develop- the implementation of eco-friendly initiatives, propel-
ment opportunities for e-commerce, it also introduces ling SMEs towards a greener and more sustainable
inherent security risks. These risks, in turn, pose new business model (Singhal et al., 2023).
Applied Data Science and Smart Systems 399
Cloud computing simplifies the expansion of global cloud computing’s capabilities in mass data storage,
reach for e-commerce businesses, enabling them high-speed computing, and resource allocation can
to effortlessly offer their products and services to a enable the creation of an e-commerce application
wider audience worldwide. This enhanced accessibil- model.
ity fosters significant sales growth, new market explo- Given the current trend towards cloud computing,
ration, and ultimately, an expanded global presence. we anticipate a growing number of e-commerce web-
Businesses can seamlessly scale their operations to sites migrating to the cloud. Our approach aims to
meet fluctuating demands and adapt to changing mar- bring both the applications and data of e-commerce
ket conditions, ensuring a smooth and uninterrupted websites closer to the clients, thereby enhancing the
customer experience. performance of cloud-based e-commerce site hosting
As technology advances, cloud computing will services. The robust storage, operational, and secu-
continue to play a pivotal role in shaping the future rity functions of cloud computing, coupled with its
of e-commerce, providing businesses with the agil- efficient resource allocation and sharing, establish a
ity, scalability, and insights necessary to navigate the solid foundation for the development of e-commerce
ever-changing digital landscape and achieve long- recommendation engines, leading to a novel business
term success. recommendation approach.
By utilizing cloud computing, e-commerce organi-
Discussion zations can significantly reduce the hardware and soft-
ware costs associated with web mining, consequently
Cloud computing has emerged as a transforma- increasing enterprise profitability. Furthermore, the
tive force in the e-commerce landscape, empow- adoption of cloud computing, which offers a pay-as-
ering businesses of all sizes to harness its power to you-use model, can substantially reduce setup and
achieve remarkable growth and success. Embracing maintenance expenses.
cloud computing solutions can pave the way for
rapid growth and long-term success for e-commerce
Future work
startups, enabling seamless adaptability to chang-
ing demands, flexible operations expansion without Security and privacy enhancements: Future research
significant infrastructure investments, and enhanced should prioritize the development of robust secu-
cloud-based security measures that safeguard cus- rity and privacy measures tailored to e-commerce
tomer data and foster consumer trust and loyalty. transactions in cloud computing environments.
E-commerce startups can access advanced analyt- To safeguard sensitive user data and ensure
ics and machine learning capabilities through cloud the confidentiality and integrity of e-commerce
computing. These tools provide valuable insights into transactions, it is crucial to implement advanced
customer behavior, optimize pricing strategies, and encryption techniques, multi-factor authentica-
deliver personalized shopping experiences. This level tion, secure data storage protocols, and robust
of customization enhances customer satisfaction and access control mechanisms.
fosters loyalty, contributing to the overall success of Scalability and performance optimization: With the
the e-commerce startup. e-commerce industry’s relentless growth, re-
They are able to effortlessly handle increased searchers should delve into strategies to bolster
website traffic, expand their product inventory, and the scalability and performance of cloud-based
efficiently manage their operations without any per- e-commerce systems. This entails a thorough
formance issues. The scalability, cost-effectiveness, examination of techniques like load balancing,
flexibility, and enhanced data security offered by resource allocation, and caching mechanisms to
cloud computing contribute significantly to their effectively manage surging user demands and en-
overall business growth and competitiveness in the sure seamless user experiences, even during peak
e-commerce industry. traffic periods.
Cost efficiency and sustainability: Enhancing the cost-
Conclusion efficiency and sustainability of e-commerce op-
erations in cloud computing necessitates future
With the exponential growth of the internet and research that explores strategies to reduce in-
commercial websites, the volume of information in frastructure costs, energy consumption, and the
e-commerce systems continues to increase rapidly. carbon footprint associated with running e-com-
E-commerce has emerged as the prevailing business merce applications in the cloud. Additionally, re-
model in contemporary society, providing numerous search can focus on implementing green comput-
shopping and consumption platforms for people. ing practices and optimizing resource utilization
Based on our research, we believe that leveraging
Applied Data Science and Smart Systems 401
to bolster both cost-efficiency and sustainability based on cloud computing. 2021 IEEE Asia-Pacific
in cloud-based e-commerce systems. Conf. Image Proc. Elec. Comp. (IPEC), 1100–1103.
Mobile commerce (m-commerce) integration: As Taherkordi, A., Feroz, Z., Yiannis, V., and Geir, H. (2018).
mobile devices increasingly permeate our daily Future cloud systems design: challenges and research
directions. IEEE Acc., 6, 74120–74150.
lives, research should focus on integrating cloud
Li, Y. and Junfeng, L. (2020). Risk management of e-com-
computing with m-commerce to enhance user
merce security in cloud computing environment. 2020
experiences and enable seamless mobile transac- 12th Int. Conf. Meas. Technol. Mechatr. Autom. (IC-
tions. This necessitates a thorough exploration MTMA), 787–790.
of techniques like mobile application develop- Li, Y., Hong, Z., and Li, Z. (2019). Research on the con-
ment, context-aware computing, and location- struction of e-commerce security risk assessment
based services to foster personalized and loca- model based on cloud computing. 2019 11th Int.
tion-specific e-commerce experiences on mobile Conf. Meas. Technol. Mechatr. Autom. (ICMTMA),
platforms. 589–592.
Big data analytics for personalization: Harnessing Baghdadi, Y. (2013). From e-commerce to social commerce:
the vast trove of data generated by e-commerce a framework to guide enabling cloud computing. J.
Theoret. Appl. Elec. Comm. Res., 8(3), 12–38.
transactions, future research can investigate the
Al-Jaberi, M., Nader, M., and Jameela, A.-J. (2015). E-com-
utilization of big data analytics in cloud comput-
merce cloud: Opportunities and challenges. 2015 Int.
ing to deliver personalized recommendations, Conf. Indus. Engg. Oper. Manag. (IEOM), 1–6.
targeted marketing, and enhanced customer ex- Krypa, A. and Anni, D. (2016). Impacts of cloud comput-
periences. This involves developing sophisticat- ing in e-commerce. INTED 2016 Proc., 1812–1819.
ed algorithms and frameworks to analyze user IATED, 2016.
behavior, preferences, and purchase history, en- Faccia, A., Corlise Liesl Le, R., and Vishal, P. (2023). In-
abling businesses to provide tailored recommen- novation and e-commerce models, the technology
dations and elevate customer satisfaction. catalysts for sustainable development: The Emirate of
Ethical and legal considerations: The burgeoning Dubai case study. Sustainability, 15(4), 3419.
realm of e-commerce through cloud computing Shi, L., Wenyong, W., Jinghui, W., and Su, Y. (2017). Re-
search on the application and development trend of
necessitates a thorough examination of ethical
cloud computing based on E-commerce. 2017 6th Int.
and legal implications, particularly in the areas
Conf. Comp. Sci. Netw. Technol. (ICCSNT), 339–342.
of data ownership, data protection, and consum- IEEE, 2017.
er rights. Future research should prioritize the Singh, J., Singh, S., Singh, S., and Singh, H. (2019). Evaluat-
development of frameworks and guidelines that ing the performance of map matching algorithms for
ensure e-commerce practices in the cloud align navigation systems: An empirical study. Spat. Inform.
with ethical principles and adhere to applicable Res., 27, 63–74.
laws and regulations. This includes delving into Wang, B. and Jian, T. (2016). The analysis of application
topics such as data privacy, consent manage- of cloud computing in e-commerce. 2016 Int. Conf.
ment, and transparency in data usage to foster Inform. Sys. Artif. Intel. (ISAI), 148–151.
trust and maintain the confidence of both busi- Hao, W., James, W., and Chris, T. (2013). Accelerating e-
commerce sites in the cloud. 2013 IEEE 10th Cons.
nesses and consumers. By addressing these criti-
Comm. Netw. Conf. (CCNC), 605–608.
cal concerns, the future of e-commerce in cloud
Rao, T. K. R. K., Sajid, A. K., Zeenat, B., and Ch Divakar.
computing can be shaped in a responsible and (2013). Mining the e-commerce cloud: A survey on
sustainable manner. emerging relationship between web mining. E-comm.
Cloud Comput. 2013 IEEE Int. Conf. Comput. Intel.
References Comput. Res., 1–4.
Xiaofeng, Y., Yumei, Z., and Yang, W. (2013). The innova-
Yang, Zaoli, Qin Li, Vincent Charles, Bing Xu, and Shi- tion of e-commerce financial service product based
vam Gupta. (2023). Online Product Decision Support on cloud computing—taking Alibaba Finance as an
Using Sentiment Analysis and Fuzzy Cloud-Based example. 2013 10th Int. Conf. Ser. Sys. Ser. Manag.,
Multi-Criteria Model Through Multiple E-Commerce 259–261.
Platforms. IEEE Transactions on Fuzzy Systems, 31, Treesinthuros, W. (2012). E-commerce transaction security
3838–3852. model based on cloud computing. 2012 IEEE 2nd Int.
Singhal, S., Laxmi, A., and Himanshu, M. (2023). Sustain- Conf. Cloud Comput. Intel. Sys., 1, 344–347.
able manufacturing integrated into cloud-based data Almarabeh, T. and Yousef Kh, M. (2019). Cloud computing
analytics for e-commerce SMEs. 2023 Int. Conf. Artif. of e-commerce. Mod. Appl. Sci., 13(1), 27–35.
Intel. Smart Comm. (AISC), 1436–1440. Liu, T. (2011). E-commerce application model based on
Xuecong, C., Li, Z., and Chen, S. Design and implemen- cloud computing. 2011 Int. Conf. Inform. Technol.
tation of e-commerce recommendation system model Comp. Engg. Manag. Sci., 1, 147–150.
52 Navigating blockchain-based clinical data sharing: An
interoperability review
Virinder Kumar Singlaa, Amardeep Singh and Gurjit Singh Bhathal
University College of Engineering, Punjabi University, Patiala, Punjab, India
Abstract
Blockchain technology holds significant promise for revolutionizing the healthcare industry by eliminating the need for
trusted third parties and enhancing data security. However, despite substantial progress, challenges such as interoperabil-
ity, performance, access control, scalability, and integration persist, hindering widespread adoption. This paper focuses on
exploring the critical issue of interoperability in healthcare systems. Clinical data, encompassing patient vitals, medical im-
ages, medications, and more, is now managed digitally through electronic medical records (EMR), electronic health records
(EHR), and personal health records (PHR). These systems, while offering convenience, are susceptible to security breaches
and data fragmentation. This paper identifies these research gaps and proposes a comprehensive solution to address block-
chain interoperability in healthcare, aiming to create an efficient, secure, and integrated healthcare data management eco-
system. The research seeks to benefit patients, healthcare providers, and the medical research community by facilitating the
seamless exchange of critical healthcare information.
Keywords: Blockchain, blockchain-based healthcare, blockchain interoperability, interoperability, clinical data, data security,
healthcare systems
vksingla@[Link]
a
Applied Data Science and Smart Systems 403
platform for managing healthcare data (Hölbl et al., research. They analyzed its evolution and various
2018). issues. They explained basic structure of a smart con-
The traditional healthcare systems depend largely tract and its working principle for blockchain archi-
on trusted third parties. Many a time these parties tecture, analyzed installation procedure across the
have proven to be trust breaching (Schmeelk, Dragos, hyperledger fabric, Ethereum and electro-optical sys-
and Debello, 2021). Blockchain technology offers a tem (EOSIO) blockchains, and produced a contrast-
potential solution to this problem as it relies on dis- ing study. They further introduced the deployment
tributed consensus against central authority in tradi- process and potential of direct acyclic graph (DAG)
tional healthcare systems. based blockchain smart contracts over Byteball,
Durneva et al. (2020) presented the use of block- InterValue and IOTA platforms. They investigated
chain technologies in healthcare. They discussed sev- the state of smart contract applications using the
eral health care applications utilizing the potential Ethereum and hyperledger fabric platforms, consid-
offered by blockchain technology. These applications ering supply chain management, Internet of Things
include: (IoT), financial transactions, and medical applica-
tions. Future research directions suggested include
• Medical information management systems (EHR issues like interoperability, integration, performance,
and EMR) privacy, formal verification and design & security
• Personal health record (PHRs) mechanism.
• Telemedicine and mHealth ElRahman and Alluhaidan (2021) presented a user-
• Data preservation system (DPS) friendly blockchain-based IoT-edge framework offer-
• Pervasive social network (PSN) ing many features to healthcare institutions such as
• Health information exchange (HIE) complete preservation of patient data, its confiden-
• Remote patient monitoring systems (RPMS) tial transmission and safe submission of the patient
• Medical research systems (MRS). examination results. However, system interoperability
still needed to be examined. Also, system performance
All these applications have revamped patient partici- under various other computational intelligence algo-
pation and control, healthcare providers’ accessibility rithms and the development of clinical decision sup-
to medical information and use of this data for medi- port system needed to be explored.
cal research. Xie et al. (2021) reported that healthcare services
can be improved by the use of blockchain technol-
Literature review ogy as it offers decentralized, immutable, transpar-
ent, and secure methods of information storage and
The use of blockchain in healthcare, like in other transport. Its integrated development with other
fields, is on the rise. Recent literature was reviewed budding technologies like AI, IoT, wearable devices,
to figure out applicability of the technology in health- cloud computing and big data, etc., could offer long-
care. Some of the literature accessed is summarized term benefits including user empowerment to exer-
below. cise better control over their health data, enabling
Adere (2022) concluded that blockchain, in health- a tamper-proof medical history and encouraging
care and IoT, is primarily used for data management better medical responsibility with ease. However,
with a prime focus on data security comprising of concerns like interoperability, efficiency, scalabil-
data-integrity, access-control and privacy-preserva- ity, security and regulatory framework were also
tion. Popular techniques used are encryption, archi- reported.
tectural designs, third-party solutions, smart contracts Fetjah et al. (2021) described a blockchain based
and authentication techniques for autonomous pro- smart healthcare system involving three-layered archi-
cessing. Also, the integration of IoT and blockchain tecture: smart medical instrument (IoT) layer, fog-
IoT, including health-IoT, is reviewed with integra- layer, and cloud-layer for remote patient monitoring.
tion mechanisms ranging from combining blockchain Data analysis was done using artificial intelligence
completely with data transfers among IoT devices (AI) and smart contracts. The proposed framework
to using it only for maintaining meta-data. Several was put to use to monitor patients with diabetes
research gaps viz., issues involving the use of a var- remotely. In addition to making proactive predictions,
ied number of smart contracts affecting the system’s anticipating future problems, and alerting a doctor
performance, data retrieval issues specifically from in the event of an emergency, the system was able to
encrypted files, and issues involving the integration of recommend treatments. Major implementation chal-
disparate healthcare systems were highlighted. lenges reported include scalability, interoperability
Lin et al. (2022) outlined the blockchain smart and limited data access control due to permissioned
contract’s operation and the state of its application blockchain used.
Applied Data Science and Smart Systems 405
Liu et al. (2021) presented a blockchain-based and bandwidth overhead, not favorable to IoT
distributed access-control mechanism for securing networks.
IoT data. It made use of the alliance chain and fog Yaqoob et al. (2021) presented various case stud-
computing concepts. On an edge node, the IoT data ies utilizing blockchain technology in healthcare in
was encrypted using the least significant bit (LSB) different countries of the world. The study included
and mixed linear and non-linear spatiotemporal Estonia’s e-health system, UAE’s national block-
chaotic systems (MLNCML) approaches. This data, chain-based platform for maintaining healthcare and
is then, further uploaded onto the cloud. Thus, solv- pharma data, Swiss hospitals using hyperledger-based
ing the issue of failed access control by providing permissioned blockchain for tracking medical devices
dynamic and fine-grained access control for IoT data. and the U.S.-based Patientory Inc.’s blockchain-based
However, further research gaps highlighted were the DApp solution facilitating health institutions to share
need for developing a lightweight consensus proto- medical data with their patients. The authors high-
col for quick confirmation and increased throughput lighted major challenges demanding research focus to
and the use of smart contracts for effective automated be scalability, regulatory framework, interoperability,
access control. potential threat issues arising out of recent advance-
Hussien et al. (2021) discussed use of blockchain ments in quantum computing, tokenization, inte-
in telecare medical information systems and e-health gration, accuracy and adoption and technical skill.
systems, reviewed and evaluated the same in terms Further future research recommendations included:
of security and privacy. The study discussed poten- the convergence of blockchain and AI, IoT-based
tial future challenges such as scalability and storage healthcare systems, integration of blockchain into leg-
capacity, blockchain size, universal interoperability acy healthcare systems, establishing blockchain legal
and standardization. Future blockchain prospects for framework, smart contracts and latency and through-
use in patient empowerment in healthcare data man- put barriers.
agement and sharing, clinical-trials, counterfeit drug Khatri et al. (2021) reported that interoperabil-
prevention, Big data, AI, 5G ultrasonic device, secu- ity, integrity, privacy, security and access control are
rity and privacy were also highlighted. the major issues of blockchain application in health-
Ejaz et al. (2021) proposed a framework, Health- care. The majority of the research covered focused
BlockEdge, with the integration of edge computing on algorithm/protocol, framework and structural
and blockchain technology. It provided friendly, secure design. Application areas for healthcare include dis-
and reliable mean for aid and remote-monitoring of tributed ledger, consensus mechanism and smart
the elderly people at home. The presented system contracts over private blockchain like Etherium and
tends to be secure, reliable, cost-effective, and resil- hyperledger framework. Also, it is reported that the
ient to network issues and offered prolonged usage major domains in healthcare using blockchain include
under diverse network issues. The proposed frame- EHR, PHR and inter & intra-institutional migration
work was compared against no blockchain system on support. Various concerns raised include security and
the parameters of power usage, delay, computing load privacy issues due to the use of personal keys, immu-
and network usage. Further future research direc- tability issues arising out of maliciously recorded
tions suggested by the author include optimization inaccurate data, scalability, interoperability and speed
using AI of collective usage of edge computing and issues.
blockchain approaches in healthcare for efficiency Newaz et al. (2020) presented an exhaustive survey
and performance improvement, developing solutions on the security and privacy issues in modern health-
for building trust among various users of to maxi- care systems. They reported that the increasing use
mally utilize the features of the blockchain in bringing of technologies like IoTs, implantable medical devices
trust between different stakeholders of multi-faceted (IMDs) and body area networks (BANs) in healthcare
distributed communication and data management not only improved the quality of patient care and
healthcare systems. treatment; but had exposed the healthcare systems
Liang and Ji (2021) reported that privacy issues to numerous cyber threats breaching their integrity,
are prevalent viz. a viz. IoT network’s nature of confidentiality, availability, privacy and security. They
scale and distribution. Blockchain has been useful listed different blockchain-based approaches, among
in overcoming various maintenance, security, data other approaches, to counter the potential challenges.
protection, and privacy & authentication issues Research directions discussed include the develop-
of IoT systems. Also, it could provide distributed ment of lightweight and symmetric cryptographic
storage, transparency, trust, and secure distributed protocols considering the emergency where communi-
IoT networks, while guaranteeing the security and cation with unauthorized personnel may be required,
privacy of the users. They reported various issues development of standard communication protocols,
such as scalability, computing complexity, latency, fault-tolerant design, intrusion detection mechanism,
406 Navigating blockchain-based clinical data sharing: An interoperability review
• Scalability of blockchain based healthcare systems Objective of the proposed research is to present
remains persistent (Fetjah et al., 2021; Khatri et a feasible solution for the problem of blockchain
al., 2021; Liang and Ji, 2021; Xie et al., 2021). interoperability for effective healthcare.
• Blockchain-based healthcare systems cannot
be seamlessly integrated with existing classical Methodology
healthcare systems (Yaqoob et al., 2021; Adere,
2022; Lin et al., 2022).
• Regulatory/legal framework for the use of block-
chain based healthcare systems; nationwide and
worldwide is not well defined (Xie et al., 2021;
Yaqoob et al., 2021).
Work proposal
Contemporary blockchain based healthcare systems
face many challenges and barriers hindering seamless
implementation of the technology. These include pri-
vacy and security issues, consensus algorithms, com-
putational power requirements, implementation costs
and integration challenges with existing healthcare
information system (Durneva et al., 2020).
Though the research is in progress to fix these issues,
almost negligible attention is drawn toward interoper-
ability aspect till present (ElRahman and Alluhaidan,
2021; Fetjah et al., 2021; Khatri et al., 2021; Xie et
al., 2021; Lin et al., 2022). A variety of blockchains
are used in healthcare systems to harness intrinsic
benefits of the technology. All such implementations
are being worked/reworked upon in isolation to fix/ Conclusion
improve any performance issues. But no work is being
carried out to make such implementations to interop- Blockchain technology is emerging as a powerful solu-
erate i.e., to communicate and exchange healthcare tion to address vulnerabilities in traditional healthcare
information, which may be vital for realizing effective systems heavily reliant on third-party intermediaries.
healthcare for mankind. Electronic medical records (EMR), electronic health
408 Navigating blockchain-based clinical data sharing: An interoperability review
records (EHR), and personal health records (PHR) Heart, T., Ben-Assuli, O., and Shabtai, I. (2017). A review
have digitized clinical data, offering greater accessi- of PHR, EMR and EHR integration: A more personal-
bility but also exposing security and privacy concerns. ized healthcare and public health policy. Health Pol-
Blockchain’s decentralized ledger and cryptographic icy Technol., 6(1), 20–25. [Link]
hlpt.2016.08.002.
features eliminate the need for central authorities,
Hölbl, M., Kompara, M., Kamišalić, A., and Zlatolas, L.
transforming data sharing in healthcare.
N. (2018). A systematic review of the use of block-
Its applications span drug supply chain manage- chain in healthcare. Symmetry, 10(10), 470. https://
ment, access control, healthcare credential manage- [Link]/10.3390/sym10100470.
ment, clinical trials, and research, giving individuals Khatri, S., Alzahrani, F. A., Md Tarique, J. A., Agrawal,
control over their records. However, challenges per- A., Kumar, R., and Ahmad Khan, R. (2021). A sys-
sist, with interoperability being a primary concern. tematic analysis on blockchain integration with
Isolated blockchain systems hinder data exchange, healthcare domain: Scope and challenges. IEEE
especially during emergencies, and performance issues Acc., 9, 84666–84687. [Link]
affect scalability and efficiency. CESS.2021.3087608.
Robust data access control, seamless integration Liang, Wenbing, and Nan Ji. (2022). Privacy challenges of
IoT-based blockchain: a systematic review. Cluster
with existing healthcare systems, and clearer regula-
Computing. 25(3): 2203–2221.
tory frameworks are necessary. Research into block-
Lin, S.-Y., Zhang, L., Li, J., Ji, L., and Sun, Y. (2022). A sur-
chain interoperability within healthcare is essential to vey of application research based on blockchain smart
enable diverse blockchain systems to freely exchange contract. Wirel. Netw., 28(2), 635–690. [Link]
data, creating a more efficient healthcare ecosystem. org/10.1007/s11276-021-02874-x.
This research aims to benefit patients, healthcare Liu, Y., Zhang, J., and Zhan, J. (2021). Privacy protection
providers, and the broader medical research commu- for fog computing and the Internet of Things data
nity by fostering secure, integrated healthcare data based on blockchain. Cluster Comput., 24(2), 1331–
management. 1345. [Link]
In conclusion, blockchain enhances healthcare Maloy, C. (2022). Library guides: Data resources in the
data security and management, but ongoing research health sciences: Clinical data. [Link]
[Link]/hsl/data/findclin.
is crucial to address challenges, advance interoper-
Newaz, Akm Iqtidar, Amit Kumar Sikder, Mohammad
ability, and create a secure and integrated healthcare
Ashiqur Rahman, and A. Selcuk Uluagac. (2021).
landscape. A survey on security and privacy issues in modern
healthcare systems: Attacks and defenses. ACM Trans-
References actions on Computing for Healthcare. 2(3): 1–44.
[Link]. (2020). Global electronic
Adere, E. M. (2022). Blockchain in healthcare and IoT: health records (EHR) market (2020 to 2025) - by
A systematic literature review. Array, 14, 100139. product, component, end-user, region, competition,
[Link] forecast & opportunities - ResearchAndMarkets.
Durneva, P., Cousins, K., and Chen, M. (2020). The current Com. May 27, 2020. [Link]
state of research, challenges, and future research direc- news/home/20200527005390/en/Global-Electronic-
tions of blockchain technology in patient care: Sys- Health-Records-EHR-Market-2020-to-2025---by-
tematic review. J. Med. Internet Res., 22(7), e18619. Product-Component-End-user-Region-Competition-
[Link] [Link].
Ejaz, M., Kumar, T., Kovacevic, I., Ylianttila, M., and Har- Schmeelk, Suzanna, Denise Dragos, and Joan Debello.
jula, E. (2021). Health-blockedge: Blockchain-edge (2021). What Can We Learn about Healthcare IT Risk
framework for reliable low-latency digital health- from HITECH? Risk Lessons Learned from the US
care applications. Sensors, 21(7), 2502. [Link] HHS OCR Breach Portal. 3993–3999.
org/10.3390/s21072502. Xie, Y., Zhang, J., Wang, H., Liu, P., Liu, S., Huo, T., Duan,
ElRahman, S. A. and Alluhaidan, A. S. (2021). Blockchain Y.-Y., Dong, Z., Lu, L., and Ye, Z. (2021). Applica-
technology and IoT-edge framework for sharing tions of blockchain in the medical field: Narrative re-
healthcare services. Soft Comput., 25(21), 13753– view. J. Med. Internet Res., 23(10), e28613. https://
13777. [Link] [Link]/10.2196/28613.
Fetjah, L., Azbeg, K., Ouchetto, O., and Andaloussi, S. J. Yaqoob, Ibrar, Khaled Salah, Raja Jayaraman, and Yousof
(2021). Towards a smart healthcare system: An ar- Al-Hammadi. (2021). Blockchain for healthcare data
chitecture based on IoT, blockchain, and fog com- management: opportunities, challenges, and future
puting. Int. J. Healthcare Inform. Sys. Informat. recommendations. Neural Computing and Applica-
(IJHISI), 16(4), 1–18. [Link] tions: 1–16.
SI.20211001.oa16.
53 Analysis of data backup and recovery strategies in the
cloud
Sumeet Kaur Sehra1,a and Amanpreet Singh2
1
Wilfrid Laurier University, Waterloo, Canada
2
Lovely Professional University, Punjab, India
Abstract
This study examines modern data backup and recovery techniques in cloud computing settings. An in-depth literature
analysis highlights the changing environment by examining the effects of various techniques. This study offers empirical
insights into strategy choices and difficulties by employing a rigorous methodology that involves data collecting from diverse
cloud service providers and enterprises. The results have demonstrated that choosing a cloud provider impacts how a plan is
implemented and perennial concerns about data security, compliance, and cost management. Further, the latest technologies,
including blockchain-based data integrity and artificial intelligence (AI)-driven anomaly detection, have also been discussed.
Keywords: Blockchain, data integrity, artificial intelligence, anomaly detection, diverse cloud service
a
sksehra@[Link]
410 Analysis of data backup and recovery strategies in the cloud
service layer, and a perceptual recognition layer. The IoT devices, which are frequently sensitive to network
central link bridging the material and informational attacks (Dalal, 2023). Focusing on Spark is a quick
worlds comprises the visual recognition layer pow- and comprehensive framework for handling massive
ered by perception technology. This layer includes amounts of data. Spark performs at speeds that are
specialized adaptive electronic devices for human 100 times faster than Hadoop and MapReduce when
data retrieval and automatic data collecting tools, memory resources are abundant. It outperforms these
e.g., radio frequency identification (RFID) and sensor rivals by a factor of 10 when leaking data to disk,
networks. even when memory is limited. Spark’s capabilities for
The network infrastructure layer’s primary respon- complex directed acyclic graphs (DAGs), intended for
sibility is to connect the Internet with lower-layer in-memory data processing, give it its performance
evaluation and recognition tools, providing access to prowess. Spark is an object-based and operational
applications at the upper layers. The Internet and the programming framework implemented in Scala,
next generation form the basis of the IoT, supported allowing for the fluid manipulation of remote datas-
by various wireless networks that provide Internet ets akin to local collection objects (Brar et al., 2022;
access and rely on robust computation and mass stor- Rahul, 2022). Its defining characteristics are rapid
age capacities for significant data collection. implementation, user-friendly operation, adaptability,
The extensive application layer reflects the chang- and compatibility (Lai, 2022).
ing environment of online applications, which is A significant question is handling the difficult data
influenced by the development of processing power backup task (Ramesh, 2023). Previously, compa-
(Iwona, 2023). Early data services focused on email nies or integrators were in charge of building these
and file transfers, but modern user-centric network systems and ensuring they complied with all speci-
applications include social networking, video stream- fications and performed at their best. The IT envi-
ing, and online gaming. Despite its increasing popu- ronment in businesses has changed over time. Still,
larity, the IoT has inherent weaknesses caused by the backup procedures frequently experienced alterations
enormous number of endpoints and difficulties asso- without a systematic methodology, omitting to con-
ciated with connectivity and collaboration among sider the importance of syncing with basic standards
Applied Data Science and Smart Systems 411
For example, let R = 7 and N = 4, then PDL = 1 research studies and industry publications show valu-
- (1-7) ^ 4 = -1295. This implies that PDL is ex- able data and insights. Official government publica-
tremely low, or it can be said that it is effectively tions and regulations provide an additional source
zero. of reliable information. The most recent techniques
Backups created from snapshots and advancements in protecting data in the cloud can
A snapshot of the cloud’s resources is taken using be found in the documentation provided by cloud
snapshot technology. It is a productive backup service providers and on technology news websites.
method without affecting system performance Online discussions, social networking sites, and spe-
(Twana, 2022). Snapshots can retrieve informa- cific groups encourage debate and user experiences,
tion and are particularly useful when you need adding practical insights to the research. Books, con-
to return to a certain condition quickly. Vendors ference proceedings, and seminars are all excellent
provide snapshot services, including AWS, Azure, resources for learning how recovering and backing
and Google Cloud. up data in the cloud is changing (Xiaojun, 2022).
Replication and redundancy Researchers can develop a comprehensive picture of
Implementing data redundancy and replication the best practices, difficulties, and emerging trends
across many cloud servers or zones of avail- in this vital topic by utilizing these secondary data
ability improves data availability and durability sources.
(Surbhi, 2015). This method stores data in sever- The secondary data included in the Scopus author-
al places to protect against calamities or outages ing project was taken from various journal articles. It
at data centers. This is made more accessible by consists of a broad range of data drawn from these
services like “Azure Geo-Replication” and “AWS publications that have been carefully collected and
S3 Cross-Region Replication”. arranged into Excel files. Many useful visualizations
Cloud-to-cloud restoration and graphic representations have been created using
Cloud-to-cloud backup methods can benefit these Excel sheets. Data collection provides insight-
businesses employing various cloud-based re- ful information on various study problems and is the
sources, such as SaaS apps. You can back up data basis for the study’s analytic approach.
using these services, such as “Veeam” or “Dru-
va,” from one cloud environment (such as Mi-
Empirical results
crosoft Office 365 or G Suite) to another cloud
(such as “AWS, Azure, or Google Cloud”). For The performance of data restoration and backup pro-
the protection of crucial corporate data kept on cedures in the cloud can be better understood through
numerous cloud platforms, it is crucial. empirical data. These findings offer quantifiable infor-
Cost-benefit analysis mation on restoration speed, recovery time, afford-
It draws insights into profitability when organi- ability, and dependability. They empower businesses
zations shift to cloud computing in each layer. to make wise judgments, improve their methods for
The three layers are base cost estimation, data cloud-based data security, and guarantee data resil-
pattern-based, and project-specific cost estima- ience in changing cloud settings.
tion. Equation 2 can be used to calculate the Figure 53.4 represents the challenge-response-
cost-benefit ratio. verification process time overhead. It is clear that
the processing overhead for challenge-response veri-
Cost – benefit Ratio = fication gradually increases as the number of chal-
Value of Data-Cost of Backup and Recovery lenge data blocks increases. Even if it exists, the rise
(2)
Cost of Backup and Recovery in the process of verification overhead is still barely
noticeable. Usually, the time cost of generating prob-
lems is far lower than that of generating responses.
Description of dataset
However, this cost gradually increases as more chal-
To get knowledge about this crucial field of IT, sec- lenging data blocks are added. However, when the
ondary data collecting for studying cloud-based number of data blocks exceeds a critical threshold,
backup and recovery of data solutions entails explor- such as 2000, the challenge’s time cost significantly
ing current sources. Researchers can access various increases and converges with the confirmation pro-
information from research databases, publications, cess’s cost.
and articles through a thorough literature review. Choosing proper algorithms is crucial for building
Whitepapers, reports, and analyses that offer valu- an effective processing system for keeping data and
able data and trends are frequently found on sector- backups within the Spark platform. The “APCA seg-
specific websites and in the documentation of cloud mentation”, “ratio R”, “differential D”, and “dura-
service providers (Surbhi, 2015). Additionally, market tionik T” techniques are all considered in this analysis.
Applied Data Science and Smart Systems 413
Figure 53.5 shows that the errors related to the other shows the comparison to information backup
three two-stage approximation approaches are sig- management.
nificantly lower than those associated with APCA It is essential to provide quick access to vital recov-
segmentation. The trials have shown that the ratio R ery data. Off-site data storage is a component of the
method outperforms the efficiency of duration-based described strategy, which calls for keeping backups
point identification and produces the lowest aver- elsewhere. Two methods are used: physically mov-
age error. As a result, when evaluating the results of ing the data and writing it into removable drives. In
Spark-based processes, the ratio R method is the best the case of a failure, it is crucial to have quick access
option for picking key points (Figure 53.6). procedures in place for adequate recovery. Figure
Events may have unfavorable effects on the IT infra- 8 shows the comparison study of off-server copy
structure and the broader business operations. Fires management.
in buildings, problems with central heating systems in The benefit of this strategy is its simplicity of orga-
server rooms, and unexpected equipment theft are a nization. The difficulties in media retrieval cause the
few examples. One successful approach is to establish requirement to move data to preservation and the
procedures for recovering data during a disaster. potential for media damage while in transit. It entails
In such circumstances, a strategy to reduce data copying data to a different location across a network
loss is to keep backup storage at a distant place, away channel. Figure 53.9 shows an example of storage
from the main server equipment area. Figure 53.7 device management.
414 Analysis of data backup and recovery strategies in the cloud
Conclusion Iwona, K., Jackowski, A., Lichota, K., Welnicki, M., Dub-
nicki, C., and Iwanicki, K. (2023). InftyDedup: Scal-
In conclusion, this study investigated cloud-based able and cost-effective cloud tiering with deduplica-
data backup and recovery techniques in depth. A tion. 21st USENIX Conf. File Stor. Technol., 23,
thorough literature study revealed these tactics’ 33–48.
expanding importance in response to rising cloud use Rahul, K. and Venkatesh, K. (2022). Centralized and de-
and related data threats. Using a strong approach, it centralized data backup approaches. Proc. Int. Conf.
gathered data from numerous cloud service providers Deep Learn. Comput. Intel. ICDCI, 2021, 687–698.
and companies and learned important lessons. doi:10.1007/978-981-16-5652-1_60.
Empirical findings showed various backup tech- Lai, Y. L., Rana, M. E., and Al Maatouk, Q. (2022). Criti-
niques, with firms choosing methods following data cal review of design considerations in forming a
cloud infrastructure for SMEs. 2022 Int. Conf. Dec.
volume, recovery goals, and budgetary restrictions.
Aid Sci. Appl. (DASA), 1537–1543. doi: 10.1109/
As enterprises prioritize data redundancy and disaster
DASA54658.2022.9765167.
recovery capabilities, the choice of cloud provider has Ramesh, G., Logeshwaran, J., and Aravindarajan, V. (2023).
emerged as a crucial aspect. Data security, compliance A secured database monitoring method to improve
and cost management were problems, but AI-driven data backup and recovery operations in cloud com-
anomaly detection and blockchain-enhanced data puting. BOHR Int. J. Comp. Sci., 2(1), 1–7. doi:
integrity were promising advancements. To success- 10.54646/bijcs.019.
fully balance accessibility, security, and cutting-edge Brar, P. S., Shah, B., Singh, J., Ali, F., and Kwak, D. (2022).
technology, it is essential to continually assess and Using modified technology acceptance model to eval-
change cloud backup and recovery procedures, as this uate the adoption of a proposed IoT-based indoor
study highlights. disaster management software tool by rescue work-
ers. Sensors, 22(5), 1866, [Link]
s22051866.
References Rehman, A. U., Agular, R. L., and Barraca, J. P. (2022).
Fault-tolerance in the scope of cloud computing.
Dalal, A. (2023). Secure cloud migration strategy (SCMS):
IEEE Acc., 10, 63422–63441. doi; 10.1109/AC-
A safe journey to the cloud. Int. Conf. Cyber Warfare
CESS.2022.3182211.
Sec., 18(1), 1–6. doi: 10.34190/iccws.18.1.1038.
Twana, H. S., Sharif, K. H., and Rashid, B. N. (2022). A
Dajun, C., Li, L., Chang, Y., and Qiao, Z. (2021). Cloud
survey of comparison different cloud database perfor-
computing storage backup and recovery strategy
mance: SQL and NoSQL. Passer J. Basic Appl. Sci.,
based on secure IoT and spark. Mob. Inform. Sys.,
4(1), 45–57. doi: 10.24271/psr.2022.301247.1104.
1–13. doi: 10.1155/2021/9505249.
Surbhi, K. (2015). A survey on dynamic load balancing
Durga, V. S. K., Fatima, Y., and Mailewa, A. B. (2022). Data
techniques in cloud computing. Adv. Comp. Sci. In-
integrity attacks in cloud computing: A review of iden-
form. Technol. (ACSIT), 2(7), 87–91.
tifying and protecting techniques. Int. J. Res. Publ.
Xiaojun, S., Huang, Y., Liu, Z., and Yang, Y. (2021). Reduc-
Rev., 3(2), 713–720. doi: 10.55248/gengpi.2022.3.2.8.
ing the service function chain backup cost over the
Kiranpreet, K., Guillemin, F., and Sailhan, F. (2022). Con-
edge and cloud by a self-adapting scheme. IEEE Trans.
tainer placement and migration strategies for cloud,
Mob. Comput., 21(8), 2994–3008. doi: 10.1109/
fog, and edge data centers: A survey. Int. J. Netw.
TMC.2020.3048885.
Manag., 32(6), e2212. doi: 10.1002/nem.2212.
54 Landslide identification using convolutional neural
network
Suvarna Vani Koneru, Harshitha Badavathula, Prasanna Vadttityaa and
Sujana Sri Kosarajub
Velagapudi Ramakrishna Siddhartha Engineering College, Andhra Pradesh, India
Abstract
Landslide identification poses a significant challenge in ensuring the safety of vulnerable regions. Accurate detection is cru-
cial for timely mitigation efforts. In this study, we propose a convolutional neural network (CNN) model based on transfer
learning for classifying landslide-prone areas using a diverse dataset. The dataset comprises satellite images of landscapes
categorized into distinct classes. Addressing class imbalance, we employ preprocessing techniques and oversampling meth-
ods. The images are resized to a standardized 32 × 32 pixel format to enhance model efficiency. The model leverages a
pre-trained CNN architecture and incorporates additional layers for fine-tuning. Training utilizes the Adam optimizer and a
suitable loss function. These strategies are vital for optimizing the model’s performance and ensuring precise classification of
landslide-prone areas. Evaluation is conducted based on accuracy metrics, showcasing the model’s proficiency in capturing
essential features of landscapes prone to landslides. Our proposed approach holds promise for geologists and environmental
experts, offering a high-accuracy solution for identifying landslide-prone regions and facilitating effective mitigation strate-
gies (MDPI, 2023)
Keywords: Landslide identification, convolutional neural networks (CNN), transfer learning, oversampling, pattern recogni-
tion
vadttityavsprasanna@[Link], bksujanasri31@[Link]
a
Applied Data Science and Smart Systems 417
non-landslide sites are randomly split into training monitoring essential soil moisture levels and ground
(70%) and testing (30%) groups to build a landslide movement, is one of the project’s key components.
inventory map. Sixteen landslip conditioning ele- Utilizing geographic information systems (GIS) tech-
ments relating to topography, hydrology, lithology, nology, this real-time data is incorporated into an
and land cover are incorporated into a GIS database. intricate geospatial framework. Algorithms for ML
The ReliefF approach is used to assess these aspects’ are used to model the intricate interactions between
importance and provide guidance for model construc- rainfall quantity, topography, and previously recorded
tion. Landslide susceptibility maps (LSMs) are then landslides, allowing the system to produce precise and
generated with the models of the NB classifier, RF timely alerts. Three things are expected to happen as
classifier, and FDEMATEL-ANP. Metrics including a result of the project: first, data-driven insights will
the area under the curve (AUC), mean absolute error increase the accuracy of landslip prediction; second,
(MAE), root mean square error (RMSE), Kappa index an intuitive and user-friendly interface will be created;
(K), and overall accuracy (OAC) are used to assess the and third, efficient channels of communication will be
effectiveness of the model. Based on the data, the RF established to inform local communities and relevant
classifier is the most promising and ideal model for authorities about alerts. To protect people and prop-
landslide susceptibility in the studied area. It has an erty in landslide-prone areas, the “SESAMO Early
elevated K and OAC value of 0.8435 and 92.2%, a Warning System for Rainfall Triggered Landslides”
low MAE (0.1238), RMSE (0.2555), and a high AUC aims to close the gap between scientific research and
value of 0.954. This thorough approach highlights the practical disaster management (Puma et al., 2015)
RF classifier’s superior performance over other mod- The “Deep Learning-Based Landslide Susceptibility
els and offers insightful information for assessing the Mapping” paper offers an original and thorough solu-
Kysuca river basin’s susceptibility to landslides. (Sahin tion to the problems associated with reliably determin-
et al., 2020; Ali et al., 2021; ResearchGate, 2023) ing and mapping landslide susceptibility in geologically
The research “Landslip detection in the Himalayas sensitive locations. Traditional approaches to landslip
using ML algorithms and U-Net” approaches the susceptibility mapping frequently entail human inter-
challenging issue of landslip hazards, which are fre- pretation and analysis of numerous geographic datas-
quent in the Himalayan region, in an original and ets, which can be time-consuming, biased, and unable
cutting-edge way. The Himalayas are known for their to capture intricate interactions between contributing
rough topography, geological instability, and suscep- elements. This project suggests using state-of-the-
tibility to a variety of natural occurrences, such as art deep learning methods into the procedure to get
landslides, which pose serious risks to communities, around these constraints. The paper depends on the
infrastructure, and the environment. This project uses collection and preparation of sizable datasets includ-
a comprehensive approach to address these issues. ing pertinent geographical characteristics. The deep
The project’s core consists of the integration of state- learning models are trained and fine-tuned using these
of-the-art technology, with a particular emphasis on datasets, allowing them to learn the underlying cor-
ML techniques and the U-Net architecture. While the relations and produce susceptibility maps with a bet-
U-Net architecture, a CNN, specializes in segmenting ter level of accuracy. greater accuracy and detail than
images – a vital duty in landslip detection – ML allows permitted by conventional approaches. The project is
computers to learn from data and make informed expected to produce high-resolution landslide suscep-
judgments. The endeavor makes use of these tools for tibility maps, which will give important insights into
analyzing a broad dataset made up of high-resolution landslide-prone regions and classify them according
satellite images, elevation data from LiDAR data, and to a gradient of vulnerability. These maps provide a
topographical details of the Himalayan environment crucial resource for decision-makers and urban plan-
(Meena et al., 2022). ners to set priorities for risk reduction initiatives, cre-
A comprehensive method to deal with the impend- ate sustainable land use plans, and create effective
ing threat of rainfall-induced landslides is offered by disaster preparedness systems (Azarafza et al., 2021).
the “SESAMO Early Warning System for Rainfall The methodology for the paper titled “Landslide
Triggered Landslides” paper. The paper seeks to Recognition by Deep CNN and Change Detection”
create a robust early warning system that improves involves a comprehensive four-step process. First, a
preparedness and lessens the impact of landslides on deep CNN is constructed and trained using datasets
sensitive regions by utilizing cutting-edge technologies derived from remotely sensed (RS) images contain-
and real-time data integration. The system aims to ing historical landslide information. This CNN serves
reliably forecast landslip events in response to shifting as the foundation for subsequent stages. Second, an
rainfall patterns by integrating meteorological data, object-oriented change detection CNN (CDCNN)
ground monitoring sensors, and geospatial analysis. with a fully connected conditional random field
The creation of a vast sensor network, capable of (CRF) is implemented, leveraging the insights gained
418 Landslide identification using convolutional neural network
from the trained CNN to detect changes indicative of images provide enlarged, high-resolution views that
landslides. The third stage involves the optimization enable detailed inspection of hilly terrain. The data-
of the preliminary CDCNN through post-processing set exhibits considerable variety in area parameters,
methods tailored to refine and enhance the accu- illumination, and image quality, which poses oppor-
racy of change detection. Finally, the results are fur- tunities and problems for the creation of reliable and
ther augmented by information extraction methods, effective classification algorithms. The study intends
including trail extraction, source point extraction, to enhance the effectiveness of landslide detection
and attribute extraction. Image block processing and and classification algorithms by utilizing this exten-
parallel processing strategies are employed through- sive dataset. This will allow for the more accurate
out to significantly improve speed, a crucial aspect and dependable identification of particular patterns,
when dealing with RS images covering extensive geo- structures, and features indicative of different hill
graphical areas. The methodology is validated using situations.
two landslide-prone sites in Hong Kong, demonstrat-
ing high speed, exceeding 80% accuracy, and practical Pre-processing and exploratory data analysis
applicability in real-world scenarios. This integrated To learn more about the dataset, exploratory data
approach showcases the effectiveness of combining analysis is done before to training the models. With a
deep learning, change detection, and information count plot, the frequency distribution of the classes is
extraction techniques for accurate and efficient land- shown, giving a brief overview of how certain features
slide recognition from RS images (Shi et al., 2021; are distributed throughout the data set. Furthermore,
ResearchGate, 2023). the way that each feature is distributed is examined to
identify patterns or correlations.
Objectives
Data augmentation and pre-processing
The main objective of this paper is to identify land- Different data augmentation techniques were used
slides and analyze those that may occur in the future to improve the generalizability and robustness of
and to distinguish between images that depict land- the model. These techniques include using random
slides and images that do not. transformations such as rotating, panning, zoom-
ing, and moving training images to create enhanced
Methodology versions of the original database. By exposing the
model to a wider range of variables, the expanded
Dataset data set improves the model’s ability to generalize
The deepglobe land classification dataset, which con- unseen information. In addition, the image is scaled
sists of a varied collection of images depicting hilly to a standard size of 32 × 32 pixels, which provides
landscapes with various land structures, is used in consistent dimensions for post-processing. The pixel
this work. This 4188-image dataset includes examples values of the reconstructed images are normalized to
of both landslide events and non-events, displaying encourage better approximation during training and
a range of hill topography with varying elevations to keep the magnification constant. This normaliza-
and slopes. Extensive attributes and metadata are tion process involves subtracting the mean and divid-
appended to every image, offering significant insights ing by the standard deviation of the database to bring
into the characteristics of hills that are essential for the pixel values to a standard range (MDPI, 2023)
classification and analysis. These satellite-captured (Figure 54.1).
Model architecture
The envisioned model architecture for landslide iden-
tification harnesses the power of CNN with transfer
learning to proficiently detect and classify pertinent
features associated with landslides. The architectural
design unfolds in the following sections.
The model commences with an input layer tailored
to accommodate RGB images of dimensions 32 ×
32 pixels. To exploit prior knowledge and optimize
performance, a pre-trained CNN model serves as
the foundational base. This base model is initialized
with weights gleaned from the “imagenet” dataset,
allowing the model to inherit valuable insights from a
Figure 54.1 Resized landslide image diverse array of image classification tasks.
Applied Data Science and Smart Systems 419
Following the transfer learning base, customized The testing set was kept solely for assessing the
convolutional layers are introduced. These layers model’s performance on hypothetical data, enabling
are meticulously crafted to fine-tune the model spe- an evaluation of its generalization skills. In order to
cifically for the task of landslide detection, capturing enable efficient training, suitable components were
intricate patterns and spatial information inherent to selected, utilizing the sparse categorical cross-entropy
landslide features. loss function to address cases involving several classes
A critical flattening operation ensures after the con- of categorization. and the effective gradient-based
volutional layers, converting the output feature maps optimization tool, the Adam optimizer. The model was
into a coherent one-dimensional vector. This stra- trained over several epochs. Additionally, to evaluate
tegic transformation ensures seamless connectivity the model’s performance on unknown data at each
with subsequent layers, facilitating the extraction of epoch and guarantee its generalization skills through-
high-level representations. Following this, the output out the training process, a validation set – typically a
consists of any of the ones among landslide detected portion of the training data was employed. The model
and landslide not detected, representing the prob- incorporates dropout, a regularization technique, to
abilities of the input image belonging to each of the minimize unnecessary risk. Every training update,
2 classes of landslide detection. The softmax activa- dropout arbitrarily eliminates a portion of the input
tion function is applied to the output layer to obtain a units to keep the model from becoming overly depen-
probability distribution across the 2 landslide classes, dent on any one characteristic and strengthening gen-
enabling effective classification. eralization capacity. The model’s performance was
assessed using an alternative set of tests. The model
Model training and evaluation is trained with image tests, and its prediction ability
In this study, the landslide classification model pro- is assessed using performance metrics like accuracy,
posed was trained and evaluated using the deep- precision, recall, and F1 scores. By measuring overall
globe land cover classification dataset. This dataset accuracy, the ratio of true positive predictions to true
consists of a diverse collection of satellite images of positive predictions, the ratio of true positive predic-
hills and mountains, encompassing different types of tions to all true positive examples, and the balanced
situations, including both landslide-occurrence and F1. score, this metric offers a thorough understand-
non-occurrence cases. To ensure reliable results, the ing of the performance of the model. Think about
dataset was partitioned into distinct training and test- recall and accuracy. Furthermore, a thorough analysis
ing subsets, enabling rigorous evaluation of the mod- of this model’s performance is conducted to obtain a
el’s performance. The training set served the purpose greater knowledge of its efficacy in accurately catego-
of optimizing the model’s parameters and capturing rizing landslides and in delivering useful information
intricate patterns within the images (Figure 54.2). for advancement and development.
Abstract
This study presents an optimized multi-level thresholding technique for the retinal vessel segmentation in fundus images. The
proposed methodology involves pre-processing steps such as illumination compensation and adaptive histogram equaliza-
tion to enhance vessel visibility. The segmentation is performed using a Tsallis-based multi-level thresholding algorithm, and
the results are evaluated using ground truth images. Performance metrics including sensitivity, specificity, and accuracy are
calculated, demonstrating the effectiveness of the proposed method in detecting blood vessels accurately. The average values
of specificity, sensitivity and are found to be higher compared to an existing method. The proposed technique also shows
promising results in differentiating normal and abnormal retinal images.
Keywords: Retinal vessel segmentation, fundus images, multi-level thresholding, Tsallis thresholding, illumination compensa-
tion, adaptive histogram equalization, sensitivity, specificity, accuracy
vishalishapar78@[Link]
a
Applied Data Science and Smart Systems 425
the active contours without edges method for image Remove objects below a certain threshold to smooth-
segmentation. en the image (e.g., optic disc and lesions).
Enhance contrast using histogram equalization tech-
Existing approaches nique.
One prevalent approach in the literature involves uti-
lizing morphological operations for vessel extraction. Step 2: Vessel segmentation
Abramoff et al. (2010) applied wavelet-transform-
assisted-morphological gradient operation, along Convert the pre-processed image to grayscale.
with the CLAHE to pre-process the low-contrast fun- Apply thresholding to convert the grayscale image in
dus images. The Morphological gray level hit as well to a binary image, where pixels above a certain
as the miss transform, incorporating multi-structur- threshold are considered vessels and the rest one
ing elements of varying orientations, was employed are background.
for blood vessel-background separation (Robinson et Use morphological operations (e.g., dilation, opening,
al., 1997). erosion) to extract structurally suitable vessel
pixels.
Proposed approach
To address the limitations of existing methods, this Step 3: Post-processing
study proposes an integrated approach combining
morphological pre-processing and threshold-based Convert the binary image back to grayscale.
segmentation. The morphological pre-processing Apply another thresholding to remove unwanted ar-
phase focuses on identifying linear vessel structures eas from the foreground.
(Dasgupta et al., 2009). The subsequent thresholding- Use morphological operations to further remove
based segmentation separates blood vessels from the noise and unwanted regions from the segmented
retinal fundus image. Further refinement involves vessels.
skeletonization to pinpoint vessel intersections,
enhancing segmentation accuracy. Step 4: Performance evaluation
Performance parameters Compare the segmented vessels with the ground truth
The research’s success will be evaluated using per- using metrics like specificity sensitivity and ac-
formance parameters such as specificity sensitivity, curacy.
and, accuracy. These metrics provide a comprehensive Assess the quality of vessel segmentation using receiv-
assessment of the proposed method’s ability to accu- er operating characteristic (ROC) analysis and
rately segmenting the blood vessels from the retinal other relevant performance measures.
fundus images.
In the following study, the proposed methodology The existing approach involves pre-processing the
will be elaborated, experimental results will be pre- input image to enhance vessel visibility, followed by
sented and analyzed, and the overall conclusions and thresholding and morphological operations for ves-
implications of the research will be discussed. sel segmentation. Post-processing steps are performed
to refine the segmented vessels and performance met-
Algorithm of existing technique rics are used to evaluate the accuracy of segmentation
against the ground truth.
Input: Retinal fundus image
Output: Segmented blood vessels
Algorithm of proposed technique
Step 1: Pre-processing
Input: Colored retinal fundus image
Extracting the green channel from RGB image, and it Output: Segmented blood vessels
provides better vessel-background contrast. Step 1: Pre-processing
Apply 2D wavelet transform to decompose the im-
age into smaller sections with different frequency Load the colored retinal fundus image from the local
components. disk.
Remove low-frequency and keep only the high-fre- Convert the RGB image to grayscale using the RGB2
quency components from the wavelet coeffi- gray function.
cients. Apply thresholding, for convert the grayscale image
Denoise the image using a Gaussian filter to ensure to a black and, white binary image based on Ot-
uniform intensity distribution. su’s method.
426 Retinal vessel segmentation using morphological operations
Smooth the cross-section points where blood vessels mitigate this issue. Figure 55.3 displays the illumina-
intersect, by enhancing ridge and bridge points. tion-corrected images of normal subjects, along with
Perform dilation to identify and remove small objects their filtered low- frequency components. Similarly,
(disks) from the image. Figure 55.4 presents the illumination-corrected
Apply erosion to fill the disk areas with the back- abnormal images and their corresponding filtered
ground color, removing unnecessary regions. components. These results demonstrate that the illu-
mination correction method enhances vessel edges
Step 2: Vessel segmentation and overall contrast, improving vessel segmentation.
Step 3: Post-processing
Illumination compensation
Uneven illumination in retinal images can hinder
the accuracy of segmentation. The cubic spline illu-
mination compensation technique was employed to Figure 55.2 Abnormal gray scale images
Applied Data Science and Smart Systems 427
Figure 55.3 Illumination corrected normal images (a) Figure 55.4 Illumination corrected abnormal images
and corresponding low frequency components (b) (a) and corresponding low frequency component (b)
Clique function
To further enhance vessel edges, the illumination-
corrected and histogram-equalized images under-
went treatment with the clique function. Figures 55.7
and 55.8 exhibits the improved edges in normal and
abnormal images, respectively. This process not only
enhanced the vessel edges but also improved the edge
pixels contributing to microvasculature information, Figure 55.5 Illumination corrected and, adaptive his-
alongside other anatomical features. togram equalized normal images
428 Retinal vessel segmentation using morphological operations
Proposed 89 93 93.2
Existing 79 89 86
Performance evaluation
The performance evaluation of proposed blood vessel Table 55.2 presents the average ratios of vessel to
segmentation method was conducted using two key vessel-free area for both the proposed and existing
metrics, true positive fraction (TPF) and false positive methods. The proposed approach exhibited higher
fraction (FPF) which were employed for comparison ratios for both normal and abnormal images, show-
with an existing approach. casing its efficacy in distinguishing between vessel and
non-vessel regions. This characteristic contributes to
Average sensitivity, specificity and accuracy its effectiveness in differentiating between normal and
Table 55.1 provides the average specificity sensitivity abnormal images.
and accuracy values for both the proposed and exist-
ing methods. The proposed optimized Tsallis multi- ROC analysis
level threshold method demonstrated higher average Receiver operating characteristic (ROC) analysis was
sensitivity (89%) and specificity (93%) compared used to evaluate the approaches diagnostic accu-
to the existing method. This superior performance racy. The ROC curves for the proposed method are
implies the proposed approach’s ability to detect depicted. The area under the curve (AUC) for the pro-
blood vessels accurately, even in challenging cases. posed method was 0.95, indicating its superior per-
Moreover, the higher specificity highlights the pro- formance in blood vessel detection. This high AUC
posed method’s efficiency in detecting relevant vascu- value confirms the proposed algorithm’s efficiency
lature, including those in abnormal images. and accuracy in detecting blood vessels.
Applied Data Science and Smart Systems 429
Abstract
Accurate liver segmentation in a computed tomography (CT) images are essential for various medical applications. This
study presents a novel liver segmentation method that combines shape prior features and a modified Chan-Vese (CV) model.
The proposed algorithm extracts shape characteristics from a training set using statistical shape modeling. A comprehensive
comparison of the proposed approach with existing methods is conducted based on performance parameters like maximum
symmetric surface distance (MSD), relative volume difference (RVD), average symmetric surface distance (ASD), root mean
square symmetric surface distance (RMSD), and volumetric overlap error (VOE). Experimental results on SLIVER and IR-
CAD datasets showcase the superiority of proposed method. The algorithm demonstrates enhanced segmentation accuracy
and efficiency, making it a valuable asset in medical image analysis.
Keywords: Liver segmentation, CT images, shape prior features, CV model, medical image analysis, performance evaluation,
statistical shape modeling, SLIVER dataset, IRCAD dataset, segmentation accuracy
veerpalsync@[Link]
a
Applied Data Science and Smart Systems 431
Level set methods stands out as a prominent approach. This model com-
Level set methods extended the capabilities of active bines active contours and level set methods to achieve
contours by enabling the evolution of curves in a robust segmentation results. An overview of the CV
higher-dimensional space. These methods allowed model and its key components are discussed in the
for better handling of complex shapes and topology following sections.
changes. However, they were computationally inten-
sive and required careful parameter tuning (Saito et Chan-Vese model
al., 2017). The Chan-Vese (CV) model is a widely used technique
for image segmentation, particularly for medical
Graph cut and region-based approaches images like liver segmentation. It is formulated as an
Graph cut and region-based methods brought about energy minimization problem that aims to find a con-
significant advancements by modeling liver segmen- tour that divides the image into regions correspond-
tation as an optimization problem. These techniques ing to the object of interest and the background. The
integrated image data with spatial information and key advantage of the CV model is its ability to handle
have shown promising results in handling shape vari- intensity in homogeneity and adapt to object shape
ations. Nevertheless, they often required extensive variations (Getreuer et al., 2012).
manual intervention and were sensitive to initializa-
tion (Kitrungrotsakul et al., 2015; Li et al., 2015). Components of the CV model
• Energy functional: The CV model defines an
Machine learning-based methods energy functional that consists of data fidelity
Machine learning techniques, including random for- and regularization terms. The data fidelity term
ests, Support Vector Machines, and, Convolutional makes sure that the evolving contour aligns with
Neural Networks, have gained traction in recent years. intensity gradients, where the regularization term
CNNs, in particular, have demonstrated remarkable encourages smoothness of the contour (Heimann
capabilities in capturing intricate features and learn- et al., 2009).
ing shape variations directly from data. However, • Level set evolution: The active contour evolves
they demand substantial computational resources and based on the minimization of the energy function-
extensive training data (Jin et al., 2017; Pawar et al., al over iterations. The evolution is achieved using
2020). partial differential equations (PDEs) that modify
the contour’s shape while adhering to the image’s
Statistical shape models (SSMs) intensity properties (Zhang et al., 2010).
Statistical shape models have emerged as a poten- • Region-based energy: The CV model employs
tial solution for handling shape variations in liver region-based energy terms, which are calculated
segmentation. These models capture the shape vari- within the evolving contour and its complement
ability within a training dataset and utilize statistical (outside the contour). These energy terms capture
measures to guide the segmentation process. They the difference in intensities between the object
offer adaptability to shape changes but may struggle and background regions (Li et al., 2015).
with unseen variations not present in the training data • Balloon force: An additional term, known as the
(Zheng et al., 2017). balloon force, is often incorporated to adjust the
contour’s shape and handle concavities or con-
Integration of shape prior information vexities in the object boundary.
One key limitation of many existing methods is their
inability to efficiently handle liver shape variations
caused by pathologies or anatomical differences. To Proposed algorithm
address this, some recent approaches have integrated The proposed algorithm seeks to improve the liver
shape prior information into the segmentation process. segmentation from CT images by enhancing the exist-
These methods leverage training datasets to learn the ing CV model with the incorporation of shape prior
liver’s expected shape variations, aiding in accurate information. This additional information helps over-
segmentation even in the presence of deformations. come some of the limitations of the CV model and
leads to more accurate and efficient liver segmenta-
Existing algorithm tion results. An overview of the proposed algorithm’s
key components and steps are discussed below.
The existing liver segmentation algorithms have
undergone continuous evolution to address the chal- Step 1: Input CT image – Begin by inputting the CT
lenges posed by shape variations, noise, and artifacts image containing the liver region that needs to be
in CT images. Among these algorithms, the CV model segmented.
432 Liver segmentation using shape prior features with Chan-Vese model
Step 2: Shape prior information – Introduce shape It consists of CT images of liver structures, allow-
prior information extracted from a training dataset, ing for validation on a different dataset assessing the
such as the SLIVER dataset. This information pro- algorithm’s generalization capability.
vides knowledge about the expected shape of the liver Volumetric overlap error (VOE): Measures the
and aids in guiding the segmentation process. overlap error between segmented and ground truth
volumes.
Step 3: Structural and statistical features extraction
Relative volume difference (RVD): Quantifies the
– Extract structural and statistical features from the
relative difference in volume between segmented and
input CT image. These features may include entropy,
ground truth regions.
homogeneity, dissimilarity, and fractal characteristics.
Average symmetric surface distance (ASD):
These features help capture important information
Measures the average distance between surfaces of
about the liver’s characteristics.
the segmented regions (Figures 56.1–56.4).
Step 4: Enhance CV model – Modify the existing CV It amply illustrates how much superior the pro-
model to incorporate the extracted shape prior infor- posed work segmentation is to the current method.
mation and the structural and statistical features. This
enhancement aims to improve initialization, conver-
gence, and accuracy of the segmentation process.
Step 5: Segmentation using enhanced model – Utilize
the enhanced CV model to segment the liver from the
CT image. The integration of shape prior information
and feature-based constraints guides the contour evo-
lution process more effectively.
Step 6: Performance evaluation – Quantitatively
assess the performance of the segmentation by cal-
culating various metrics such as relative volume dif-
ference (RVD), volumetric overlap error (VOE), root
mean square symmetric surface distance (RMSD),
average symmetric surface distance (ASD) and maxi-
mum symmetric surface distance (MSD).
Step 7: Comparison with existing model – Compare
the segmentation outcomes produced by the proposed
algorithm with those obtained using the traditional
CV model. Evaluate the improvement in terms of
Figure 56.1 The comparison of existing and the pro-
accuracy, robustness, and efficiency.
posed methods. (a) Represents the original images. (b)
Represents the ground truth images. (c) Represents
Dataset and parameters the existing method and (d) Represents the proposed
method.
The proposed algorithm for the liver segmentation
and, enhancement of the CV model is evaluated using
two well-known datasets: the SLIVER dataset and the
IRCAD dataset.
SLIVER dataset
This dataset provides a comprehensive collection of
CT images containing liver structures. It serves as the
training dataset for shape prior information extraction.
The SLIVER dataset includes a variety of liver
images with different shapes, sizes, and pathological
conditions, making it suitable for robust algorithm
training.
IRCAD dataset
The IRCAD dataset is used as the testing dataset for Figure 56.2 Shows the original image belongs to the
evaluating the performance of the enhanced algorithm. SLIVER dataset
Applied Data Science and Smart Systems 433
Abstract
The significance of video analytics in online conferencing has grown quite a lot due to the drastic increase in the usage of
online conferencing platforms in recent years. Deep learning techniques have shown potential across various applications
within video analytics, including tasks such as object detection, scene classification, and event recognition. This review paper
takes a look at a comprehensive overview of the research on the use of deep learning for video analytics in the context of
online conferencing. The paper summarizes the various approaches used for video analytics, including convolutional neural
networks, various machine learning models, and other techniques, and evaluates their performance on a range of datasets
and usecases. The review also shows the challenges and limitations associated with these methods, including scalability
and variability in accuracy, and the need for further research in these areas. The paper concludes by discussing the future
potential of deep learning for video analytics on online conferencing and the potential for new and innovative approaches
to emerge in the future.
Keywords: Video analytics, video conference analytics, deep learning, neural network, transformers, convolution neural
network
average is utilized to accumulate evidence over time, Review based on convolution neural network
aiding in making the final prediction. This research In the publication titles “A video shot boundary detec-
paper introduces a deep learning-based methodology tion utilizing CNN features”, a novel method for
for conducting real-time sentiment analysis on video CNN feature extraction is proposed for video shot
streams. The approach focuses on classifying the emo- border detection. By concurrently extracting features
tional expressions of the subjects over time by lever- from video sequences using a CNN model on a GPU,
aging visual and/or audio information present in the the suggested method streamlines the expression of
data stream. In their research, the authors utilized the video and shortens computation time for shot identi-
RAVDESS (Ryerson audio-video database of emotional fication. In order to improve shot detection recall and
speech) dataset. They achieved an accuracy of 90.74%, precision, the approach also accounts for local frame
surpassing the baselines that range from 11.11% to similarity and dual-threshold sliding window similar-
31.48%. The deep learning model proposed based on ity. The results of the experiments demonstrate that
multiple modalities shows promising results for real- the proposed strategy performs better in terms of F1
time sentiment analysis of video streams, but scalabil- score and speed than alternative models. Only three
ity is an open problem (Yakaew et al., 2021). videos were used to train the model, so it may process
The paper titled “Highlight detection with pairwise information more slowly than other approaches. In
deep ranking for first-person video summarization” the future, it could be beneficial to include motion fea-
introduces a novel approach utilizing a pairwise deep tures or different kinds of gradual transition features
ranking model to identify highlights in first-person in order to improve the accuracy (Liang et al., 2017).
video (FPV). The primary objective is to generate a The research paper titled “Detection and recog-
comprehensive summarization of the video content. nizing cursive text from video images” introduces a
The model employs deep learning techniques to learn comprehensive framework for detecting and recog-
the similarity or highlight and non-highlight segment nizing italic text in video images, specifically focus-
relationship of videos and uses a two-stream network ing on Urdu text. Textual content within videos holds
structure to represent segments of videos from com- significant relevance for various applications such as
plementary information on the appearance of video semantic search, alert generation, and advanced tasks
frames and temporal dynamics across frames. The like opinion mining and content summarization. The
model is evaluated against state-of-the-art methods study incorporates several techniques, including CNN
– RankSVM, and demonstrates a notable accuracy and DNN-based object detection, as well as LSTM
improvement of 10.5% when applied to 100 hours of networks. The authors utilize the ICDAR dataset for
first-person video (FPV) spanning 15 distinct sports training and evaluating the system. The proposed
categories. Additionally, the model exhibits superior framework demonstrates exceptional performance in
summary quality, as evidenced by a user study involv- recognizing and identifying Urdu italic text, achieving
ing 35 human subjects. The model proposed is cat- an impressive recognition F-measure of 88.3% and a
egory-independent and incorporates both temporal recognition rate of 87%. The paper not only presents
and spatial information, but the weakness is that it the framework for detecting and recognizing textual
may not generalize well to other types of videos and content in video images but also introduces a bench-
it’s not clear how well the model would perform in mark dataset comprising over 13,000 video frames
real-world applications as the study is conducted in containing italicized text. The primary objective of
controlled settings (Yao et al., 2016). this research is to provide a valuable resource for
The paper “Superintendence video summariza- high-level applications, including real-time semantic
tion” presents a new approach for video summariza- video searching, alerting, opinion mining, and content
tion in the field of surveillance. The authors look at summarization (Mirza et al., 2010).
past research on video summarization algorithms and
datasets to create their own solution. Their approach Review based on LSTM
involves picking out key frames from the video based The paper titled “Online video summarization:
on two things: each object should be in the frame, and predicting future to better summarize present”
the objects should look good and be close together. introduces a supervised learning approach called
The authors believe their solution improves video Merry-GoRoundNet for online video summariza-
surveillance by cutting out unimportant scenes and tion. The proposed method considers both spatial
highlighting important events. The paper is helpful for and temporal relationships between video frames.
understanding the issues in video surveillance and the By employing an encoder-decoder architecture and
solutions people use. However, it doesn’t mention the convolutional LSTM, MerryGoRoundNet establishes
specific dataset used for their solution, which might spatiotemporal connections and generates summa-
impact the solutions’ accuracy and strong nature ries in an online manner. The network incorporates
(Chavan et al., 2020). unsupervised next frame prediction and supervised
Applied Data Science and Smart Systems 437
scene start detection tasks, along with a proposed structure of the video sequence. The model is evalu-
loss function that balances continuity and diversity ated using their own dataset and results are encour-
within the summary. Evaluations conducted on vari- aging. The model is however not fully compact. This
ous datasets demonstrate the superior performance paper appears to be based on machine learning tech-
of MerryGoRoundNet The method ranks favorably niques, specifically clustering algorithms. It may also
among online summarization techniques and demon- involve some elements of computer vision and natural
strates competitive performance when compared to language processing, as it uses visual and audio data
offline approaches. This approach is characterized by to analyze video content. The specific implementa-
its time and memory efficiency in comparison to non- tion of the method and the techniques used to extract
autoregressive methods, the production of diverse and organize the important video information may
and well-defined summaries, and prevention of model involve some elements of deep learning, but it is not
overfitting. The addressed challenge revolves around explicitly mentioned in the paper (Chen, 2010).
automatically generating video summaries, which is The research paper titled “Auto-summarization
particularly difficult due to the subjective nature of of audio-video presentations” addresses the need
the task (Lal et al., 2019). for efficient examination of vast amounts of online
The paper “Unsupervised video summarization with multimedia content by proposing video summaries,
adversarial LSTM networks” presents a generative which are condensed versions of the original material
architecture for unsupervised video summarization composed of key segments. The authors explore three
that combines variational recurrent auto-encoders techniques for automatically generating these summa-
(VAE) and generative adversarial networks (GAN). ries for online audio-video presentations, incorporat-
The architecture includes a summarizer network and ing information from the audio signal, slide transition
a discriminator network, both of which are LSTMs. points, and previous user access patterns. In addition,
The summarizer network is responsible for choosing the paper includes the results of a user study compar-
a subset of key frames that effectively represent the ing the computer-generated summaries to summaries
input video. On the other hand, the discriminator created by the authors themselves. The study reveals
network plays a role in distin-guishing between the that users were able to acquire knowledge from the
original video and its reconstructed version gener- computer-generated summaries, but perceived them
ated by the summarizer network. The entire model is as less coherent in comparison. Notably, the com-
trained in an adversarial manner. The method is eval- puter-generated summaries were only 20–25% of
uated on four benchmark datasets (SumMe, TVSum, the length of the full presentations. Consequently, the
OVP, and YouTube) and exhibits competitive perfor- study concludes that participants expressed a prefer-
mance when compared to supervised state-of-the-art ence for using the author-generated summaries due to
methods outperforming the state of the art in video their superior coherency. One downside of the sug-
summarization by 2–5%. The paper does not men- gested method is that the computer-generated sum-
tion any specific weaknesses or limitations of the pro- maries are not as well organized as the ones made by
posed method. The problem addressed in this context humans. Additionally, the computer-generated sum-
is unsupervised video summarization, which involves maries are much shorter in duration (He et al., 1999).
the task of selecting a subset (sparse) of video frames The paper titled “Creating summaries from user
that effectively represent the content of the input videos” introduces a fresh approach to making video
video (Mahasseni et al., 2017). summaries and establishes a new standard for user
videos containing numerous interesting events. The
Approaches based on machine learning method uses “superframe” segmentation to break
This paper “Video presentation board: A semantic down the video into segments. It then gauges the
visualization of video sequence presents video presen- visual interest for each superframe based on various
tation board”, a new video summarization method low-, medium-, and high-level features. By doing this
that visualizes sequence of video in a static image assessment, the approach identifies the most informa-
for the purpose of efficient representation and quick tive and captivating subset of superframes to gener-
overview. This method uses a new video shot cluster- ate a comprehensive and engaging video summary. To
ing technique that utilizes both visual and audio data assess its performance, the method uses benchmark
to analyze video content and collect important shot data, including multiple human-generated summary
information. The authors propose a multi-level video data collected through controlled psychological
summarization method that abstracts both locations experiments. This objective evaluation allows a thor-
and interested objects and characters, and then orga- ough assessment of summarization methods and
nize and synthesize a suitable amount of selected video provides valuable insights into video summarization
information using special visual languages according techniques. The results of this evaluation show that
to the relations between video events and the temporal the proposed method achieves high-quality results
438 Online video conference analytics: A systematic review
actions, thereby enhancing the overall quality of sum- educational and news videos, but its performance on
marization (Lu et al., 2013). different types of videos and real-world scenarios is
The paper “Text extraction in video images” pro- not specified in the paper (Guo et al., 2016).
poses a method for extracting text information from The paper “Text extraction in video” presents a
video sequences. It involves examining the frequency comprehensive system designed for the detection,
of high horizontal energy in a video frame and per- localization, extraction, tracking, and binarization of
forming structural operations to remove the back- text in general-purpose videos. The method employs a
ground. The method uses temporal information and multi-faceted approach that combines edge detection
DCT coefficients to evaluate the energy and filter out and various other techniques to achieve accurate text
non-text blocks. The proposed method uses its own detection. The system is capable of processing differ-
dataset and is found to be more efficient than Wang ent types of video formats, including JPEG images,
et al.’s method, with better results on images with MPEG-1 bit streams, and live video feeds. The study
complex backgrounds. The weakness of the proposed observed that no single algorithm could effectively
method is not mentioned, but it addresses the problem detect all forms of text, necessitating the use of a
of text extraction from video sequences. The conclu- cascaded set of constraints to address this issue. By
sion is that the method is effective and efficient for text employing these constraints, the method achieved
extraction from video sequences (Yen et al., 2008). strong performance with high accuracy and a low
The paper “Text extraction from video images” rate of false alarms. However, it should be noted that
presents a new method for extracting text from video detecting text in low-contrast backgrounds still poses
frames, specifically video data from the Malayalam a challenge, indicating an area for further improve-
news channel “Mathrubhumi News.” The method ment in the system. Overall, the paper highlights the
extracts 13 different features, by employing both effectiveness of the proposed system in extracting text
spatial and frequency domain features, the algorithm from videos, showcasing its capabilities in various
aims to classify whether an image contains text or video formats and its ability to minimize false alarms.
not. The validation of the algorithm involves the It also acknowledges the need for continued research
application of classification techniques such as Simple to enhance text detection in challenging scenarios,
Logistic, J48, and random forest, with an average such as low-contrast backgrounds (Yen et al., 2008).
success rate of 98%. The strength of the proposed In this paper the author proposes a method for
method is that it extracts relevant features specific video summarization based on the analysis of video
to Malayalam scripts, leading to improved accuracy structures and highlights. It uses a normalized cut algo-
and speed. However, a weakness is that the method rithm for scene modeling and motion attention mod-
has only been tested on one specific news channel and eling for highlight detection. The resulting temporal
may not generalize well to other types of videos or graph representation encapsulates both the structure
other languages. The results show that simple logis- and attention information of the video, but only con-
tic, a neural network-based classification algorithm, siders motion information and not other multimedia
gave the best accuracy when compared to random information. The paper’s strengths are the automatic
forest and J48, which are both tree-based classifi- detection of scene changes and generation of sum-
ers. The extracted text could be used in the future to maries, while the weaknesses are limited multimedia
assist visually impaired persons by converting it into information consideration. Future research focuses on
a sound signal (Raju et al., 2017). improving the video attention model and developing
Text recognition and extraction from video pres- automatic video editing techniques (Ngo et al., 2005).
ents a technique for extracting text from videos and The paper “Automated whiteboard lecture video
converting it into editable form. The main focus is on summarization by content region detection and rep-
educational and news videos. The system processes resentation” presents a framework for summarizing
the video input and generates an editable text file as whiteboard lecture videos by detecting key content
output, saving time and human efforts. The strengths and keyframes using a bounding box detection
of the proposed system include its automation of the approach. The authors of the paper employ both deep
manual process of extracting text. No specific weak- learning metric and gradient feature approaches histo-
nesses are mentioned. The open problem addressed by gram to address their research problem. Through their
the paper is the difficulty of storing useful informa- experimentation and analysis, they observe that the
tion from videos in an editable form. Possible future histogram of gradient approach outperforms the deep
enhancements include allowing the user to select a metric learning approach in terms of performance and
specific portion of the screen for text extraction and effectiveness. Additionally, the authors introduce an
allowing the user to provide a video URL instead of efficient spatiotemporal graph-based tracking scheme
the video itself. The proposed system has potential as in their methodology. This scheme allows for effective
a useful tool for extracting useful information from tracking of objects and structures with-in the video,
440 Online video conference analytics: A systematic review
aiding in the segmentation process. Furthermore, they text extraction in complex video scenes, which has
propose a weighted conflict minimization scheme, been a challenging area of research in recent years.
which helps in generating keyframe summaries by This method combines multi-frame corner matching
minimizing conflicts and maximizing the coherence and heuristic rules to effectively address the chal-
and quality of the summary. The evaluation results lenges associated with Harris corner filtration in
revealed that their method achieved performance com- complex video scenes, ultimately leading to enhanced
parable to state-of-the-art techniques in terms of recall, detection accuracy through the fusion of information
f-measure, and the average of summary keyframes, from multiple frames.
as demonstrated using the Access Math dataset. The Additionally, local texture description is utilized
authors intend to conduct more in-depth exploration to assess similarity through the application of SVM.
of deep metric approaches, integrate lecturer action Experimental results based on 395-frame video
detection and text detection techniques, and extend images of four different types demonstrate the meth-
the application of these methods to other handwritten od’s effectiveness when compared to five existing text
lecture datasets. Nonetheless, there are outstanding extraction techniques (Guo et al., 2016).
challenges that must be addressed to further enhance
the approach’s performance (Kota et al., 2020). Approaches based on OCR
The paper “Hierarchical model for long-length The paper, “A video text extraction method for char-
video summarization with adversarily enhanced audio/ acter recognition”, presents a method for precisely
visual features” presents a novel method for summa- extracting only the video character portions from a
rizing long videos. The approach incorporates audio video text rectangle region to make a readable image
and visual features and adopts a hierarchical structure for OCR. The proposed method addresses the limi-
that captures temporal dependencies at both short- and tations of conventional methods which use a fixed
long-term levels within the video. The extracted fea- threshold for binarization and are not effective in
tures are refined using adversarial networks to enhance complex backgrounds with various intensities. The
deep feature extraction. The method was evaluated on proposed method focuses on extracting high-inten-
a dataset of 28 baseball videos, each with an accompa- sity regions with-in video text regions and expand-
nying editorial summary video, and produced quality ing them to encompass the entire character regions.
summaries. However, further evaluation on other types Experimental results demonstrate the superiority
of videos and benchmark datasets is needed to establish of the proposed method compared to conventional
the generalizability of the method (Lee et al., 2020). approaches. However, the result depends on the kind
The research paper titled ILS-SUMM: Iterated local of news video used, and some results have many
search for unsupervised video summarization” intro- errors due to the OCR being sensitive to noise and
duces a novel algorithm for unsupervised video sum- binarized images. Open problems include applying
marization. The algorithm utilizes the iterated local the method to other types of videos. In conclusion, the
search optimization framework to efficiently identify proposed method is a novel and effective approach
a subset of shots that accurately capture the essence for precisely segmenting character regions from com-
and meaning of the original video. The approach plex backgrounds in videos for OCR data entry (Hori
aims to minimize the overall distance between shots et al., 1999).
while adhering to a constraint on the duration of the
generated summary. Experimental evaluations per- Summary literature survey
formed on video summarization datasets clearly dem-
onstrate the superior performance of the ILS-SUMM Limitations, strengths and open problems
algorithm when compared to other existing methods. Tables 57.1 and 57.2 gives a gist of all the papers
The results showcase improved total distance metrics, referred. The authors have used various methods for
indicating the effectiveness of the algorithm. Notably, computer vision and multimedia analysis. The meth-
the paper emphasizes the scalability of ILS-SUMM ods used include DL, CNN, F1 Score, VAE, GAN,
when applied to lengthy video datasets. LSTM, hybrid of keyword spotting system, simple
However, it should be noted that the paper’s evalu- logistic, J48, random forest, multi-pronged approach,
ation is limited in terms of the available information normalized cut algorithm, bounding box detection,
regarding the number of datasets utilized and the spe- deep learning metrics, and gradient feature histogram
cific metrics employed for performance assessment approach. The datasets used range from own data-
(Shemer et al., 2021). sets, SumMe, TVSum, SST-5, ICDAR, 100 hours of
first-person videos, audio-video presentations, 10M
Approaches based on SVM hash tagged Instagram videos, and Malayalam news
The paper titled “A method of effective text extraction channel. The performance of these methods var-
for complex video scenes” introduces an approach for ies, with some showing high accuracy (98%) while
Applied Data Science and Smart Systems 441
Table 57.1 Summary of state-of-the-art methods
Tejal Chavan [3] New approach Not specified Improved video surveillance Implementation of the solution
for video monitoring by summarizing is not mentioned, which
summarization the video could affect its accuracy and
robustness
Yi-Lin Sung [4] Anchor-based SumMe The findings demonstrated Not specified
attention RNN and TVSum that the suggested approach
(ABA-RNN) datasets exhibited a competitive
nature
Daniele Comi [5] Fine-tuned SST-5, SQuAD Z-BERT-A outperforms The authors plan to further
transformers existing baselines zero-shot explore the potential of
model settings for both known Z-BERT-A in future work
intent classification and
unseen intent discovery
Rui Liang [6] CNN, F1 score Trained with 3 Method out-performed Trained using 3 videos only, low
own videos other models in terms of F1 processing speed, accuracy may
score and speed vary
Ali Mirza [7] CNN, DNN- ICDAR dataset 88.3% F measure for ICDAR only contains Urdu
based object for Urdu text text detection, 87% for text, and its ability with other
detection, LSTM recognition rate of Urdu languages is unclear
text
Shamit Lal [8] MerryGoR u-ndNet, Superior among online Not specified
LSTM summarization approaches
and competitive among
offline summarization
approaches
Behrooz VAE, GAN, SumMe, Demonstrates competitive Not specified
Mahasseni [9] LSTM TVSum, OVP, performance to supervised
YouTube SoA approaches,
outperforming one of the
best video summarization by
2–5%
Tao Chen [10] Video Own dataset Not specified Not specified
presentation
board
Liwei He [11] Audio signal Audio-video Users are able to learn Computer-generated summaries
information, presentations from computer-generated are less coherent compared to
Slide transition summaries but find them less author-generated summaries.
points, Access coherent Participants preferred to use the
patterns of author-generated summary
previous users
Michael Gygli Video Multiple High-quality results Not specified
[12] summarization human-created comparable to a manual
using summaries method
Super- frame from a
segmentation psychological
experiment
442 Online video conference analytics: A systematic review
Hansol Lee [24] Hierarchic model l 28 baseball Produced quality summaries Further evaluation on other
with adversarially videos with types of videos and benchmark
enhanced audio/ accompanying datasets is needed to establish
visual features editorial generalizability of the method
summaries
Yair Shemer ILS-SUMM Video The ILS-SUMM algorithm The paper has limited
[25] summarization outperforms other video evaluation and it is unclear
dataset(s) summarization approaches how many datasets were used
and provides solutions with and what metrics were used to
a better total distance. The evaluate performance
algorithm is scalable on a
long video dataset
Zhe Guo [26] SVM, multi-frame Four 395 Demonstrated the Not specified
corner matching, frame video effectiveness of the proposed
heuristic rules image types method when compared to
five existing text extraction
techniques
Osamu Hori OCR, high- News videos Superior to conventional The method is only tested on
[27] intensity regions methods news videos
extraction and
expansion
Atitaya Yakaew [1] Scalability Real-time sentiment analysis better than Scalability
baselines of 11.11–31.48%
Ting Yao [2] May not generalize well to 10.5% increase in accuracy and better Model may not
other types of videos and summarization of real human testing generalize well to
real-time performance is not other types of videos
proven
Tejal Chavan [3] Implementation of the Improved monitoring by removing idle Not specified
solution is not mentioned, scenes and highlighting important events
which could affect its
accuracy and robustness
Yi-Lin Sung [4] Not specified The findings demonstrated that the Not specified
suggested approach exhibited a
competitive nature
Daniele Comi [5] The authors plan to further Effectiveness of Z-BERT-A for unknown Not specified
explore the potential of intent detection outperformed existing
Z-BERT-A in future work methods. Need for further research to
evaluate the performance of Z-BERT-A
on other datasets and compare it to other
SoA models
Rui Liang [6] Trained using 3 videos CNN feature extraction outperforms Since only 3 videos
only, low processing speed, other models in terms of F1 score and are used to train it is
accuracy may vary speed difficult to say how
the model performs
in real life
Ali Mirza [7] ICDAR only contains Urdu Proposed framework achieves a high Not specified
text and its ability with other F-measure of 88.3%. The generalizability
languages is unclear of the framework to other languages and
scripts remains unclear
444 Online video conference analytics: A systematic review
Shamit Lal [8] Not specified MerryGORound Net exhibits superior Challenge of
performance among online summarization automatically
approaches and competitive performance generating the
in offline scenarios summary of a video
due to its subjective
nature remains an
open problem
Behrooz Mahasseni Not specified Combining VAE and GAN demonstrates Open problem is
[9] competitive performance when compared unsupervised video
to supervised approaches, outperforming summarization
the SoA in video summarization by 2–5%
Tao Chen [10] Not specified Method uses a new shot clustering method Model is not fully
that utilizes both visual and audio data to compact, and it
analyze video content is unclear how it
performs on large-
scale video datasets
Liwei He [11] Computer-generated Paper presents three techniques for Not specified
summaries are less coherent automatic creation of video summaries
compared to author-generated using different types of information. The
summaries. Participants auto-generated results might not be as
preferred to use the author- accurate as a human- generated result
generated summary
Michael Gygli [12] Not specified High-quality results Not sure how well
it will perform with
other data
Aditya Khosla [13] User generated videos of poor Approach to summarizing user-generated Not specified
quality might affect the result videos using web images as a prior.
Approach may not generalize to well-
produced videos
Bo Xiong [14] Limited to user generated Scalable unsupervised solution for Integrate several
videos poor quality. May highlight detection uses video duration as pre-trained domain-
not generalize well to well- an implicit supervision signal. Enhances specific highlight
produced videos. Reliance on the current state-of-the-art in unsupervised detectors to analyze
weakly-labeled annotations highlight detection test videos from
like hashtags for training novel domains
Akriti Ahuja [15] Information loss due to Approach that fuses sentiment analysis Not specified
cutting of video clips and on text, audio, and video modalities and
difficulty in handling cases provides a broader viewpoint. Increasing
based on ethnicity, voice the speed and accuracy of the system,
modulation, and accent creating better databases, creating a user-
friendly version of the multinodal system
Vignesh Rad Not specified Outperforms other traditional classifiers Testing the method
hakrishnan [16] and provides a single integrated system for used in the paper on
audio and text processing real-world videos
Nidhin Raju [19] Trained on only one domain Trained on specific script improved Not specified
and language and might not accuracy and speed in text extraction from
work well on other domains video frames
or languages
Shwu-Huey Yen Detecting text in low-contrast The system can detect, localize, extract, Not specified
[21] backgrounds remains a track, and binarize text from a general-
challenge purpose video. Detecting text in low-
contrast backgrounds remains a challenge
Chong-Wah Ngo Model only considers motion Method provides automatic detection of Model only
[22] information and not other scene changes and generation of video considers motion
multimedia information summaries information and not
other multimedia
information
SOA Methods Limitations Strengths Open problems
Osamu Hori [27] Novel and effective approach Not specified Not specified
for precisely segmenting
character regions. Applying
the method to other types of
videos is an open problem
446 Online video conference analytics: A systematic review
others outperforming state-of-the-art approaches. text on the whiteboard, exploring the use of object
However, limitations of the methods include scal- detection and tracking techniques for following the
ability, accuracy may vary, reliance on weakly-labeled teacher’s writing and highlighting important points,
annotations, and difficulty in handling cases based on developing and evaluating different approaches for
ethnicity, voice modulation, and accent. summarizing the content of the lecture and the white-
board, and incorporating user preferences and con-
Possible scope of research text into the summarization process. Additionally,
research can also be conducted on exploring the
The field of video summarization with multiple users impact of various factors such as lighting conditions,
talking at the same time presents several exciting camera angle, and writing style on the performance of
areas for research. This can include developing and handwriting recognition and summarization models,
improving automatic speech recognition models for as well as developing and comparing different evalua-
transcribing multiple over-lapping speech, exploring tion metrics for lecture summarization. Furthermore,
the use of audio separation techniques to separate there is potential for exploring the use of transfer
individual speech streams, developing and evaluat- learning and federated learning for video summariza-
ing different approaches for summarizing the content tion, as well as integrating video summarization with
of multiple concurrent speech streams, and incorpo- other multimedia processing tasks such as keyword
rating user preferences into the summarization pro- spotting and speaker diarization.
cess. Additionally, research can also be conducted
on exploring the impact of various factors such as
Conclusion
audio quality and speaker characteristics on the per-
formance of speech recognition and summarization To summarize, the domain of video analytics in online
models, as well as developing and comparing dif- conferencing utilizing deep learning offers abundant
ferent evaluation metrics for video summarization. prospects for research and advancement. The authors
Furthermore, there is potential for exploring the use of the examined papers have employed diverse deep
of transfer learning and federated learning for video learning techniques, including CNN, VAE, GAN,
summarization, as well as integrating video summa- LSTM, and hybrid keyword spotting systems, to ana-
rization with other multimedia processing tasks such lyze video content across various scenarios such as
as speaker diarization, sentiment analysis, emotion audio-video presentations, first-person videos, and
detection and keyword spotting. hash tagged Instagram videos. The outcomes of these
The field of intent and entity recognition using investigations have consistently showcased the capa-
transformers offers a vast range of research opportu- bility of deep learning methods to attain remarkable
nities. This can encompass areas such as developing accuracy, surpassing current state-of-the-art method-
and fine-tuning transformer models for joint intent ologies in certain cases.
and entity recognition, improving the accuracy and While the results of deep learning techniques for
robustness of models in noisy or real-world scenar- video analytics in online conferencing are promis-
ios, exploring the use of attention mechanisms to ing, it is essential to acknowledge their limitations.
better capture the context and relationships between Scalability remains a challenge, as the accuracy
intents and entities, and incorporating transfer learn- of these methods can vary depending on the size
ing for cross-lingual and cross-domain recognition. and complexity of the analyzed data. The reliance
Additionally, research can also be conducted on on weakly-labeled annotations can also lead to
developing and comparing different approaches for decreased accuracy, especially when the annotations
combining intent and entity recognition, such as do not adequately represent the underlying data.
using a two-stage process or a joint end-to-end model, Moreover, difficulties in handling factors like eth-
as well as integrating intent and entity recognition nicity, voice modulation, and accent emphasize the
with other NLP tasks such as sentiment analysis need for further research to ensure the robustness
and summarization. Furthermore, there is potential and effectiveness of these methods across diverse use
for exploring the use of federated learning for joint cases.
intent and entity recognition, as well as developing Despite these challenges, the application of deep
new evaluation metrics for assessing the performance learning in video analytics for online conferencing
of these models. holds significant potential for enhancing the user
In a video where a lecturer or a person is using a experience and facilitating effective analysis of exten-
white board to write information presents several sive multimedia data. Consequently, the authors antic-
exciting areas for research. This can include develop- ipate ongoing growth and development in this field,
ing and improving computer vision techniques for with the possibility of new and innovative approaches
accurately recognizing and transcribing handwritten emerging in the future.
Applied Data Science and Smart Systems 447
India
UG Student, School of Computer Science and Engineering, VIT-AP University, Amaravati, Andhra Pradesh, India
2,3,4,5
Abstract
This research study attempts to conduct a complete Coca-Cola sales analysis using data mining techniques to extract impor-
tant insights from historical sales data and consumer information. The major purpose in the competitive consumer products
business is to discover opportunities for improving sales efficiency and enabling informed decision-making for strategic
initiatives. For a comprehensive analysis, the study employs a wide dataset comprising product sales, geographic informa-
tion, customer demographics, and relevant variables, as well as external elements such as economic indicators and consumer
trends. Rigorous pre-processing techniques, such as data cleaning, integration, transformation, compression and pattern
generation are utilized to assure data quality and consistency. These methods deal with errors, deal with missing numbers,
and normalize the data for robust analysis. Sales data analysis employs various data mining techniques, including classifica-
tion, association rule mining (ARM), similarity analysis, and predictive models like decision trees and regression, to identify
pertinent patterns and insights. The results of this investigation offer Coca-Cola important insights, such as cross-selling
opportunities, high-potential market segment identification, and precise sales forecasts via predictive modeling. The results
of this study have noteworthy consequences for customized marketing approaches, interdisciplinary cooperation, and the
requirement for ongoing evaluation and modification to guarantee long-term prosperity in the sector.
Keywords: Sales analysis, data mining, Coca-Cola, consumer goods, sales growth, customer information, strategic initiatives
[Link]@[Link]
a
Applied Data Science and Smart Systems 449
marketing strategies and product performance within customer surveys for tailored insights. Employed pre-
a sample of Nigerian enterprises (Li and Wu, 2012). dictive modeling techniques like regression analysis
The study looked at how strategies affect product and decision trees for forecasting future sales trends.
sales (Yee, 2018; Saxena and Vikram, 2021). Kusrini
(2015) explored attributes for predicting buyer Association rule mining
behavior and purchase performance. This included Utilized algorithms like Apriori or FP-growth to unveil
applying classification techniques such as the ID3 patterns and associations between products, aiding in
algorithm, C4.5 algorithm, and decision trees (Julia identifying cross-selling opportunities and informing
and Peter, 2016; Vilata et al., 2010). Researchers targeted marketing strategies (Ibrahim, 2020).
Arthi and Kirubakaran (2017) and Malar and Deva
Priya (2018) examined a retail sales dataset using Cluster analysis
the WEKA interface. They assessed cluster forma- Grouped customers based on purchasing behavior,
tion correctness and compared incorrectness percent- demographics, or geographic location, enabling the
ages among four algorithms, including the standard identification of distinct market segments for tailored
K-means algorithm. The study revealed varying levels marketing and product offerings.
of cluster correctness.
Cross-validation delve into the concept and appli- Customer review
cations of cross-validation, a technique used to assess Employed a survey with ten focused questions to
the performance of predictive models, ensuring their understand consumer behaviors and preferences in
generalizability and reliability (Isa, 2019). Cluster the Guntur area, providing valuable insights for pre-
analysis discusses how data points are grouped into dictive analysis and customer review evaluation.
clusters based on similarities or patterns within the
data (Aldenderfer and Blashfield, 1984). Vilata et al. Predictive modeling
(2010) and SPS and Sivabalakrishnan, (2020) has pro- Applied regression analysis, decision trees, or machine
posed to creating predictive models to project retail learning algorithms that use market variables, cus-
sales trends based on historical data, “Predictive anal- tomer data, and past sales data to predict future sales
ysis of retail sales forecasting using machine learn- trends for demand forecasting and resource planning.
ing techniques” is carried out (Devi and Anto, 2021;
Alsayed and Çağla, 2020). “A clustering method Evaluation and validation
based on K-means algorithm” describes a particular Ensured reliability by evaluating data mining mod-
clustering technique that makes use of the K-means els through dataset splitting and performance met-
algorithm and talks about how effective it is at clas- rics. Conducted cross-validation techniques to assess
sifying data according to patterns or similarities. model robustness and generalizability. This compre-
Ibrahim (2020) proposed rare item prediction which hensive approach provides actionable recommenda-
will play a vital role in efficient mining process and tions for efficient sales growth based on data-driven
alternative methods for frequent item set generation. insights.
followed steps akin to (Li et al., 2012), ensuring accu- to cross-selling strategies and the evaluation of con-
racy in cluster determination. The application of sumption trends.
K-means clustering, guided by the optimal cluster
number identified through the elbow method, yielded Association rule mining
results consistent with referenced literature (Ziauddin The methods applied in this study, particularly uti-
et al., 2012), providing valuable insights into con- lizing the Apriori algorithm, resonate with findings
sumer behavior within specific market segments. presented in the study by (SPS and Sivabalakrishnan,
2020). The research on association rule mining algo-
(1) rithms provides a comparative analysis, offering a sim-
ilar foundation for discovering product associations
that lead to cross-selling opportunities. This align-
Count the number of clusters that is optimal (k). ment validates the study’s methodology in leverag-
The elbow approach is one strategy you can use to ing ARM to consider significant product associations
determine the ideal amount of clusters. Put K-means and further supports the approach used to encourage
clustering to use: Apply k-means clustering to the data additional purchases within a single transaction.
using the number of clusters that you have chosen and
output is shown in Figure 58.2. Load the dataset.
Convert the quantity columns to binary values (0 for
Load the dataset. 0, 1 for any positive value).
Select the relevant columns for clustering. Generate rules based on Apriori algorithm.
Extract the data for clustering. Generate frequent rules based on support measure.
Standardize the data using StandardScaler(). Apply confidence measure to generate association
Establish the most suitable number of clusters by em- rules.
ploying the elbow method for analysis.
Illustrate the optimal number of clusters graphically Compound annual growth rate
through the visualization of the elbow curve. In the context of the compound annual growth rate
Based on the elbow plot, choose the optimal number (CAGR) calculation, the research offers insights into
of clusters. association rule mining applications. Although not
Perform k-means clustering with the chosen number directly addressing CAGR, the principles of associa-
of clusters. tion rule mining methodology have been instrumen-
Add the cluster labels to the original dataset. tal in aligning with the approach of grouping data
Print the cluster centers. by year and evaluating consumption trends for soft
Display the dataset with cluster labels. drink brands. This connection substantiates the meth-
(Optional) Save the clustered dataset to a new CSV odology’s credibility in tracking and evaluating con-
file. sumption trends over time. Figures 58.2 and 58.3
represents annual growth rate.
Extract insights: Examine the characteristics of
each cluster to understand the consumption patterns Group by year to calculate total consumption per
of the area within each cluster. year.
Plot yearly consumption trends for each soft drink
Identify the characteristics of interest. brand.
Calculate the characteristics for each cluster. Calculate CAGR for each brand.
Compare the characteristics between clusters. Print CAGR for each brand.
Interpret the results.
The outcomes presented in Figures 58.3 and 58.4,
Association rule showcasing the yearly consumption trends for soft
Uncovering cross-selling opportunities drink brands and the CAGR, further substantiate the
The exploration of association rule mining tech- parallels between the findings in this study and the
niques, as well as the computation of compound methodologies detailed in prior research. From the
annual growth rate (CAGR), in the context of ana- findings, the result is like Monster has the highest
lyzing Coca-Cola’s sales data aligns with established CAGR at 29.468%, followed by Limca at 1.648%.
research methodologies documented in previous Most brands have negative CAGR, indicating that
studies (Alsayed and Çağla, 2020). These studies their sales are declining. This comparison pro-
have contributed significantly to the field of associa- vides a valuable context for understanding the cur-
tion rule mining and offer critical insights applicable rent research’s contribution within the established
452 Sales analysis: Coca-Cola sales analysis using data mining techniques for predictions
Feature selection.
Splitting into train and test sets.
Model selection and training.
Model evaluation.
MSE = mean
Predict preferred product based on encoded likeli-
hood.
Figure 58.4 Output of CAGR
After implementing the above algorithm, we get the
output in graph format (Figure 58.5), which shows
literature on association rule mining and trend analy- the likelihood to recommend and preferred product.
sis in the sales domain. This approach was rooted in established method-
ologies and was influenced by prior works (Cluster
Customer review validation Analysis Case Study by Alsayed and Çağla, 2020;
After conducting the predictive analysis for customer Singh et al., 2020; Umargono et al., 2020) providing
reviews, mean squared error (MSE) is used as valida- fundamental insights into the application of MSE in
tion metric of the proposed model. Since it provides a predictive analysis and its significance as a valida-
quantitative measure of the model’s accuracy. tion metric. The current research’s adoption of this
methodology aligns with best practices and adds
Load the dataset. to the body of knowledge established in the field
Pre-process data and handle missing values if needed. of predictive analytics and model evaluation. The
For demonstration purposes, let’s handle missing. model’s mean squared error of 0.75 suggests rela-
Perform label encoding for the categorical feature. tively low prediction error, indicating a reasonably
EDA. good fit. The predicted preferred product Maaza,
Visualize the relationship between encoded. Limca, Sprite requires further examination due to
Applied Data Science and Smart Systems 453
Figure 58.6 Yearly consumption trends for soft drink brands (with future predictions). Dotted line shows the future
sales of soft drinks.
potential label encoding or prediction interpretation market trends, and external factors to develop pre-
issues. cise sales forecasting models for Coca-Cola. These
models aligned with established methodologies in
Predictive modeling for sales forecasting predictive analytics, empowered the company to
Our research employed advanced predictive mod- optimize inventory management, enhance resource
eling techniques, integrating historical sales data, allocation, and efficiently prepare for demand
454 Sales analysis: Coca-Cola sales analysis using data mining techniques for predictions
fluctuations. The strategic insights led to notable Calculate CAGR for each brand.
improvements in supply chain efficiency, mitigat- Predict future consumption for each brand [2024,
ing stock-outs, and reducing unnecessary inven- 2025, 2026, 2027].
tory costs. Our methodology aligns closely with Combine historical and future consumption data.
prior works in predictive analytics, emphasizing the Plot consumption trends for each soft drink brand.
innovation and insights brought forth within the Print future predictions for each brand.
broader domain of sales forecasting and predictive
analysis. After implementing the above algorithm, we get
the output in graph format (Figure 58.6), which
Load the dataset. shows the yearly consumption trends for soft drink
Group by year to calculate total consumption per brands.
year.
Figure 58.7 Seasonal analysis for soft drink brands with respective to their region
Applied Data Science and Smart Systems 455
Seasonal and regional sales patterns fine-tuning of marketing strategies to align with the
Our meticulous analysis revealed significant seasonal target audience. Our findings parallel an existing
and regional sales patterns within Coca-Cola’s opera- study on marketing strategies and product perfor-
tions, enabling strategic tailoring of marketing tactics. mance in Nigeria, affirming the significance of under-
These insights equipped Coca-Cola to capitalize on standing these dynamics and reinforcing the relevance
peak demand periods and address challenges during of our research in the domain of marketing strategy
off-peak seasons, enhancing regional marketing cam- evaluation and product performance analysis.
paigns and product assortments. Aligned with estab-
lished studies on time-series forecasting of seasonal Load the dataset
item sales and evaluation of marketing strategies, our Clean column names by removing leading/trailing
research contributes valuable insights to the under- whitespaces and converting to lowercase
standing and utilization of seasonal and regional sales Calculate total sales for each product
patterns in the context of marketing strategies and Calculate year-over-year growth rates for total sales
regional sales analysis. Analyzing the impact of marketing campaigns on
sales
Load the dataset. Create a dictionary with marketing impact for each
Convert the “Season” column to numerical represen- year.
tation. Assign the adjusted sales with marketing impact
Group the data and calculate average sales percent- Plotting total sales and growth rate
age.
Visualization – Create a 5 × 3 grid for area plots. After implementing the above algorithm, we get the
output in graph format (Figure 58.8), which shows
After implementing the above algorithm, we get the the total sales and sales growth.
output in graph format (Figure 58.7), which shows
the seasonal analysis for soft drink brands. Model testing
To assess and validate our data mining models, we uti-
Product performance and market response lized the Coco-Cola dataset, dividing it into training and
Our study comprehensively analyzed historical sales testing subsets. Model training and performance evalua-
data and marketing initiatives, evaluating individual tion involved specific metrics, supported by cross-valida-
Coca-Cola product performance and the impact of tion techniques (k=5 subsets) for insights into robustness
marketing campaigns on sales. This scrutiny provided and generalizability. Our evaluation techniques align
valuable insights into customer preferences, enabling with existing research on cross-validation and the use
strategic optimization of the product portfolio, and of MSE as a metric, reinforcing the significance and
Figure 58.8 [Left] Total sales over the year [Right] year-over-year sales growth rate
456 Sales analysis: Coca-Cola sales analysis using data mining techniques for predictions
Abstract
Virtual influencers (VI), a novel nexus of technology and customer behavior, have significantly increased in relevance in the
marketing environment. It is crucial to investigate how users perceive and engage with these digital entities as we navigate a
world where the line between reality and virtuality is becoming more and more hazy. This research uses insights from social
platform data in conjunction with a thorough examination of the existing literature to present a nuanced analysis of virtual
influencers. It explores their evolution, promise, and related difficulties within the framework of computer science and infor-
mation technology. The study aims to identify the interaction between several variables, including perceptions of consumer
about the usefulness of virtual influencers, their usability, and their effects on commercial involvement in the metaverse. A
data-driven approach is taken to fill the knowledge gap in how people accept and interact with virtual influencers, gathering
survey data via a questionnaire and carrying out a rigorous analysis using a multiple linear regression model. This research
highlights the importance of virtual influencers as technology integration in modern marketing strategies in addition to elu-
cidating the concept of the metaverse perspective.
Keywords: Virtual influencer, metaverse, consumer perception, attitude analysis, digital marketing, technology integration,
virtual reality experience
[Link]@[Link]
a
Applied Data Science and Smart Systems 459
marketing efforts (Thakur et al., 2021; Choudhry et versus real-world avatar faces were processed, real-
al., 2022). As a result, researchers have been more world faces were more likely to elicit favorable emo-
interested in how users and consumers view these tions. The authors explored this about perceived trust
influencers in the context of social media. Although and approach intention, findings shows that these two
brands are most frequently using Instagram to begin are positively correlated to the emotions (Sokolova
influencer campaigns, it could be estimated that influ- and Kefi, 2020).
encer marketing and advertising may become more It may pose several ethical questions when looking
popular in the metaverse in the next years. Real and at this strong human resemblance, which highlights
virtual influencers are the two categories that exist on the need for empirical research into this phenomenon.
social media. Questions regarding the ideals they represent, their
Research in this area reflected that even though there accountability, and who is responsible for their activi-
is a similarity between human and human-like design ties are raised because some virtual influencers do not
for virtual avatars, these similarities did not always identify themselves as artificial (Porra et al., 2020).
equate to greater perceived trust (Mathur et al., 2020; Also, deep fakes appear to be so much real that people
Nissen and Jahn, 2021). According to Lou and Yuan started showing trust towards this misinformation,
(2019), Moustakas et al. (2020), and Ozdemir et al. which eventually may lead to manipulating or force
(2023), real social media influencers are those users of customers to change their decision (Etienne, 2021), as
the media platform who engages other users on daily shown by earlier research from Nightingale and Farid
basis by sharing their regular activities, opinions, and (2022). It might be challenging to determine who is
experiences and thereby establish the credibility in responsible in these situations and how this will even-
those specific products or industries. Virtual influenc- tually affect overall online trust. Instagram is specifi-
ers are artificially created beings that mimic the physi- cally utilized as a medium for influencer marketing,
cal traits and body language of people (Ozdemir et which modifies consumer perceptions (Sokolova and
al., 2023). They are developed by the integration of Kefi, 2020).
3D modeling with artificial intelligence (AI), and they Therefore, this research aims to examine and fill
are frequently designed to react to specific contexts the knowledge gap in how people accept and interact
and stimuli (Baudier et al., 2023). This scenario is fre- with virtual influencers and further analyze it using a
quently supported by the uncanny valley effect, which multiple linear regression model. The below sections
argues that up until a certain tipping point, trust and describes the research work related to the develop-
positive perception of agents rise with human like- ment of the metaverse, virtual influence and their fol-
ness until they drastically diminish and enter a val- lower’s perception towards them and overall impact
ley (Mathur et al., 2020; Mori et al., 2012). Ratings on the brand endorsed by them. Following this sec-
of trustworthiness and positive perception are only tion, it explains the methodology adopted, partici-
believed to rise after human likeness is impossible to pants profile, procedure and measures taken and
differentiate from actual humans (Mori et al., 2012; its analysis. Finally it discusses the statistical model
Mathur et al., 2020). This effect demonstrates that outcome and the work is concluded at the end of the
when consumers rate a virtual human’s perceived manuscript.
untrustworthiness as being high, they often rate posi-
tive affect and perceived trust as being low. Literature review
According to Kolo and Haumer (2018), the research
was done to analyze the influence of social media. It The virtual influencers’ domain has received very lim-
mainly focuses on how influencers create connections ited in-depth research intentions (Zhao et al., 2022);
and drive their followers towards the business they instead, the majority of recent studies have focused
are promoting. Furthermore, it is unclear whether on real influencers (Casaló et al., 2020; Haenlein et
users will be able to distinguish virtual influencers al., 2020; Farivar et al., 2022). The marketing and
from actual human influencers in pictures posted advertising through these influencers has been the
because some of them do not identify themselves as subject of numerous researches (Yew et al., 2018;
such on Instagram. Additionally, it is unclear whether Singh et al., 2020). Through the use of content mar-
virtual influencers will continue to be subject to the keting techniques on social media, influencers help
negative impacts of perceived higher unnaturalness brands by piquing the attention of their followers and
and reduced trust. According to recent studies by customers (Haenlein et al., 2020). Content produc-
Jacobson and Harrison (2022), influencers’ creation ers who actively spread material on particular sub-
and promotion of brand content is the main focus jects are known as influencers (Kim and Kim, 2021).
area. The credibility of the source theory is also a line Some qualities of the influencers, such as how many
of investigation adopted by researchers. According people follow them, how frequently and creatively
to one study that evaluated how computer-generated they engage with the people, and how well they have
460 Statistical analysis of consumer attitudes towards virtual influencers in the metaverse
collaborated with the similar or related domain influ- Nowadays mostly all brands are inclined towards
encers, have been taken into consideration in business. collaborating with virtual, and some businesses have
For instance, one study discovered a negative corre- made this their main focus. Many companies have
lation between an influencer’s involvement with their decided to introduce virtual influencers in place of real
followers and their number of followers and post- human influencers and analyze their customers liking
ings. They range from having no notoriety at all to towards them. Businesses have the freedom and the
being celebrities or experts in a particular field (Evans ability to customize their influencers following their
et al., 2017). They might share social media posts futuristic aspiration by utilizing virtual influencers.
about events they attended that were sponsored by Although the focus of these results is on the actions
brands. Additionally, they might promote services or and consequences of real Instagram influencers, a
product to raise awareness of the brand (Boerman et growing number of digitally created influencers have
al., 2017). According to a different study, influencers emerged in recent years. Considering this as a new
who are experts in their field (such as sports, fashion, adoption in technology, researchers are more focused
or beauty/cosmetics) typically have better engage- in understanding how customers perceive them as
ment for relevant product categories; influencers in compared to the real ones (Sands et al., 2022). Virtual
the beauty and cosmetics industries have the highest influencers could also benefit the industries economi-
engagement for product posts (Rutter et al., 2021). cally over period as getting real influencer on board
According to research, a high amount of self-dis- impose a huge amount of financial burden on the
closure enhances the influencer’s perceived relatabil- industry and sometimes availability is another major
ity and even friendship (Leite and Baptista 2022). concern. Overall it requires specialized and organized
The value of an influencer’s content increases with its efforts and long-term relationship (Tan and Liew,
personalization(Leite and Baptista, 2022; Ahn et al., 2020; Arsenyan and Mirowska, 2021).
2023). According to additional research, compelling Visually appealing virtual influencers have been
storytelling in Instagram posts and tales promotes built to combine certain identifying qualities of
the development of parasocial connections, which are their intended audience. Based on current research,
much more useful for encouraging purchase inten- it appears that virtual influencers are seen as being
tions (Farivar and Wang, 2022; Farivar et al., 2021). much less reliable than real-world influencers (Sands
Influencers’ content is valued by their followers more et al., 2022). Lil Miquela, 19-year-old girl from Los
when they demonstrate their own identities in it Angeles for instance, is a prominent Instagram user
(Farivar et al., 2022). Influencers divulge details about and virtual fashion influencer and has millions of fol-
their personal life, passions, occupations and view- lowers (Drenten and Brooks, 2020). She has proved
points. Influencers also have the advantage of coming that virtual influencers could influence the targeted
out as more sincere and real than superstars, espe- customers and bring value to the business with the
cially when they work with brands. Companies are skill of effective storytelling (Sands et al., 2022; Block
working with influencers more frequently as a result and Lovegrove, 2021). Humans are the ones who
of their perceived authenticity (Lee and Johnson, design and animate virtual influences. They combine
2022; Kim et al., 2021). AI with human inputs. They are virtual agents who
Influencers are creators of content who are also have taken physical form, according to (Tan and
open to working with companies and making money Liew, 2020), which is a suitable definition. Mostly the
from their online activities (Borchers, 2023). They experiences provided to the customer through real
engage the targeted customers on their online plat- human and virtual are similar and customer could
form to promote the brand presence, also they not differentiate between them which presents ethical
may conduct some offline events in different cities concerns (Porra et al., 2020). According to one study,
to bring awareness about the brand and its overall deep fake photos may be evaluated even more highly
growth (Campbell and Farrell, 2020). According to for perceived trust than images of genuine people
(Lou and Yuan, 2019), a brand’s customer interest (Nightingale and Farid, 2022).
and their perceptions of the like-minded influencer However, the bulk of studies examining users’
is very important, hence brand should focus on find- responses to virtual influencers observed that
ing the appropriate influencers for collaboration. although people are interested in understanding
Similarity enhances a follower’s sense of affiliation all facts related to virtual influencers yet they find
with the influencer. If followers form parasocial them very eerie, which lowers their perceived trust
interactions with influencers and feel a strong sense (Arsenyan and Mirowska, 2021). Based on these
of identity, they will bond with them more deeply. findings, several experts have urged for research to
Followers perceive influencers with greater credibil- understand the user perception of virtual influencer
ity as those who promote products related to their and to strategies whether collaborating with them for
areas of expertise. marketing would be beneficial for the businesses or
Applied Data Science and Smart Systems 461
influencers and their acceptance. For PEU, respon- is more than 30, there are multiple relations in the
dents were asked to measure their ease of interac- variables. All of the values in Table 59.2 are under 30.
tion with metaverse and then virtual influencers. In Therefore, there is no multi-collinearity between the
order to measure the attitude towards this technology, variables.
respondents were asked to rate the influencer in terms The multiple linear regression model’s summary
of information, discovering new products, creative is shown in Table 59.3. When evaluating the valid-
content, useful advice, and authenticity. Respondents ity of a dependent variable prediction, one metric to
were also asked about the influence on their buying evaluate is the multiple correlation coefficient, or “R”
experience with the exposure to virtual influencers. value. As mentioned in Table 59.3, R-value (0.626)
indicates a good level of prediction. The coefficient of
Analysis determination (R2) is 0.392, which is the proportion
Multiple linear regression analysis is employed to fur- of variance in the dependent variable that is explained
ther analyze the data collected through a question- by the independent variables. The 39.2% variability
naire to determine the relationship between all the in the dependent variable may be explained by the
constructs to identify the perception and acceptance independent variables, PU, PEU, and ATT.
gap about the virtual influencer. The following model The F-ratio in Table 59.4 suggests that the regres-
was developed, which was tested further to check the sion model as a whole fit the data well. The statistics
impact of PU, PEU and ATT on consumers’ BI con- shown in Table 59.4, F(3, 110) = 23.680, p <0.0005,
cerning virtual influencers. indicate that the independent factors significantly pre-
dict the dependent variable statistically.
Regression model: BI=a+[Link]+[Link]+b3. The results for coefficients are depicted in
Table 59.5. The multiple relationships between vari-
ATT+e.
ables exist if VIF is equal to or more than 10 (O’Brien,
2007; Uyanık and Güler, 2013). As shown in Table
Assumption testing: Firstly the data were analyzed
59.5, all the values for VIF are lower than 10, which
for its appropriateness for applying the multiple lin-
show no multiple linearity in the variables. The
ear regression test. The data were tested for univariate
normality assumption, for which skewness and kurto-
sis were identified (Table 59.1).
It is inferred from Table 59.1 below that the values
for skewness for data are within the acceptable range
i.e. ±, whereas one of the variable’s values for kurto-
sis are not in the acceptable range. One of the vari-
ables has its value (PU=1.202) above one, but kurtosis
coefficients do not differ greatly from the normal. The
normality assumption can also be tested by the chart
shown in Figure 59.2.
Figure 59.2 scatterplot shows that almost all the
scatterplots are in elliptic shape. There are no outliers.
In continuation to the assumption test, the data is also
tested for VIF (Table 59.5) and condition index.
The level of multi-collinearity in a regression design
matrix is indicated by a condition index. As suggested
by Uyanık and Güler (2013), if the condition index
Figure 59.2 Matrix scatterplot
PU -0.046 0.226 1.202 0.449 1 3.918 1.000 0.00 0.00 0.00 0.00
PEU 0.029 0.226 -0.513 0.449 2 0.052 8.640 0.03 0.01 0.06 0.93
ATT -0.779 0.226 0.222 0.449 3 0.016 15.887 0.97 0.15 0.29 0.00
BI -0.594 0.226 0.593 0.449 4 0.014 16.933 0.00 0.83 0.65 0.06
Applied Data Science and Smart Systems 463
Table 59.3 Model summary contributing to the prediction of BIU. The findings are
presented in the following section.
Model R R square Adjusted Std. error Durbin-
R square of the Watson
estimate Discussion
1 0.626 0.392 0.376 0.57569 1.681 Using the statistical package SPSS-23 software, an
optimal statistical model for regression was created
in the aforementioned section to represent the behav-
ioral intentions to employ virtual influencers. It has
Table 59.4 ANOVA
been found that PU and PEU are the two key vari-
Model Sum of df Mean F Sig. ables that significantly influence the behavioral inten-
squares square tions of using virtual influencers. Since the predictor
variable, attitude towards technology was not signifi-
Regression 23.544 3 7.848 23.680 0.05 cant, it had to be eliminated from the equation model
Residual 36.456 110 0.331 afterwards in the study. The results of the study show
Total 60.000 113 that behavioral intentions to use virtual influencer
technology are significantly influenced by perceived
usefulness and perceived ease of use. Participants
general form of the equation to predict the behavioral consistently showed that their intentions to embrace
intentions to use virtual influencers, from perceived virtual influencer technology were directly influenced
usefulness, perceived ease of use, and attitude towards by how valuable they thought the technology was.
the technology is predicted and obtained from the This implies that people are more inclined to use
coefficient Table. this technology if they believe it will help them meet
their needs or improve their experiences. Perceived
ease of use also turned out to be a significant fac-
BI = .566 + .236 * PU + .544 * + PEU + .051 *
tor influencing behavioral intentions. According to
ATT + e the research, people are more likely to employ vir-
tual influencer technology if they think it is natural
In Table 59.5, value of attitude towards the technol- and easy to use. This highlights how crucial it is to
ogy is not significant and retaining variables that do create virtual influencer platforms with an emphasis
not show statistical significance may cause the preci- on clarity and user-friendliness in order to promote
sion of the model to decrease. Therefore, the predictor widespread acceptance. Based on technology adop-
variables’ perceived usefulness and perceived ease of tion theories, human-computer interaction and social
use have been used in order the create the prediction media research has historically examined custom-
of outcome variable behavioral intentions to use. The ers’ behavioral intentions to connect with online or
coefficients were generated again by removing the AI-based agents, emphasizing their perceived ease of
variable ATT and the revised equation is as follows: use and utility (Moriuchi, 2019; Jhawar et al., 2023).
It’s noteworthy to point out that the study found
BI = .633 + .270 * PU + .546 * PEU + e no significant correlation between behavioral inten-
tions of using virtual influencer technology and the
However, the revised equations show very little attitude towards this technology. Although user
variations in the value of perceived usefulness and behavior has historically been greatly influenced by
perceived ease of use. Both the predictors significantly attitudes towards technology, the lack of significant
association, in this case, suggests that other character- presence of virtual influencers. International Journal
istics, such as perceived usefulness and perceived ease of Human-Computer Studies, 155: 102694, https://
of use, maybe more significant in predicting behav- [Link]/10.1016/[Link].2021.102694.
ioral intentions. Individuals who engage with technol- Asquith, K. and Fraser, E. M. (2020). A critical analysis of
attempts to regulate native advertising and influencer
ogy less frequently could also find it challenging to
marketing. Int. J. Comm., 14, 21. [Link]
develop a favorable liking towards virtual influencers
[Link]/ijoc/article/view/16123.
and their nuance. Additionally, the study conducted Baudier, Patricia, Elodie de Boissieu, and Marie-Hélène
by Ozdemir et al. (2023), also confirmed that people Duchemin. (2023). Source credibility and emotions
view virtual influencers as less reliable than their real- generated by robot and human influencers: the per-
world counterparts. Consequently, their ability to cul- ception of luxury brand representatives. Techno-
tivate a favorable brand attitude is weaker than that of logical Forecasting and Social Change, 187: 122255,
human influencers (Ozdemir et al., 2023). Essentially, [Link]
brands hoping to encourage the adoption of virtual Bente, G., Dratsch, T., Kaspar, K., Häßler, T., Bungard, O.,
influencer technology may have more success if they and Al-Issa, A. (2014). Cultures of trust: Effects of
modify their approaches to prioritize functionality avatar faces and reputation scores on German and
Arab players in an online trust-game. PLOS ONE,
and user-friendly design as opposed to depending
9(6), e98297. [Link]
exclusively on a shift in consumer perception of the
PONE.0098297.
technology. Furthermore, companies ought to utilize Block, E. and Lovegrove, R. (2021). Discordant storytelling,
well-known digital platforms as this is essential for ‘Honest Fakery’, identity peddling: How uncanny CGI
fostering an emotional bond with future generation characters are jamming public relations and influencer
(Chiu and Ho, 2023). practices. Pub. Relat. Inq., 10(3), 265–293. [Link]
org/10.1177/2046147X211026936.
Boerman, S. C., Willemsen, L. M., and Van Der Aa, E. P.
Conclusion
(2017). ‘This post is sponsored’: Effects of sponsor-
The study provides useful information to participants ship disclosure on Persuasion knowledge and elec-
who wish to drive innovation and shape the direc- tronic word of mouth in the context of Facebook. J.
tion of use of virtual influencer technology in the Interac. Market., 38, 82–92. [Link]
quickly evolving metaverse, where virtual experiences INTMAR.2016.12.002.
Borchers, Nils S. (2023). To Eat the Cake and Have It,
and interactions are progressively becoming a part of
too: How Marketers Control Influencer Conduct
daily life. A rising customer base may result from this
within a Paradigm of Letting Go. Social Media+
constant flow of information, which is advantageous Society. 9(2): 20563051231167336, [Link]
to business partners who may work with influenc- org/10.1177/20563051231167336.
ers whose fan bases align with their target market to Campbell, C. and Farrell, J. R. (2020). More than meets the
target particular demographics. Meeting user expec- eye: The functional components underlying influencer
tations and resolving particular usability concerns marketing. Busin. Horiz., 63(4), 469–479. [Link]
will enable virtual influencers to be more seamlessly org/10.1016/[Link].2020.03.003.
integrated into the metaverse and establish new chan- Casaló, L. V., Flavián, C., and Ibáñez-Sánchez, S. (2020). In-
nels for engagement and connection. As we navigate fluencers on Instagram: Antecedents and consequenc-
the ever-changing metaverse, developers and brands es of opinion leadership. J. Busin. Res., 117, 510–519.
[Link]
should prioritize strategies that enhance the perceived
Chiu, Candy Lim, and Han-Chiang Ho. (2023). Impact of
usefulness and usability of virtual influencer tech-
celebrity, Micro-Celebrity, and virtual influencers on
nology. Developing experiences that are immersive, Chinese gen Z’s purchase intention through social me-
flawless, and driven by value seems to be the key to dia. SAGE Open. 13(1): 21582440231164034, https://
encouraging broad adoption, and that makes virtual [Link]/10.1177/21582440231164034.
influencers technology a distinct and well-executed Choudhry, Abhinav, Jinda Han, Xiaoyu Xu, and Yun Huang.
strategy to engage your target audience in an infor- (2022). "I Felt a Little Crazy Following a'Doll'" Inves-
mative and enjoyable way. tigating Real Influence of Virtual Influencers on Their
Followers. Proceedings of the ACM on human-com-
puter interaction. 6, GROUP: 1–28.
References Davis, F. D. (1989). Perceived usefulness, perceived ease of
Ahn, S. J., Kim, J., and Kim, J. (2023). The future of ad- use, and user acceptance of information technology.
vertising research in virtual, augmented, and extended MIS Quart. Manag. Inform. Sys., 13(3), 319–339.
realities. Int. J. Adver., 42(1), 162–170. [Link] [Link]
/10.1080/02650487.2022.2137316. Drenten, J. and Brooks, G. (2020). Celebrity 2.0: Lil Miquela
Arsenyan, Jbid, and Agata Mirowska. (2021). Almost hu- and the rise of a virtual star system. Fem. Media Stud.,
man? A comparative case study on the social media 20(8), 1319–1323. [Link]
.2020.1830927.
Applied Data Science and Smart Systems 465
Etienne, H. (2021). The future of online trust (and why Lee, S. S. and Johnson, B. K. (2022). Are they being au-
deepfake is advancing it). AI Eth., 1(4), 553–562. thentic? The effects of self-disclosure and message sid-
[Link] edness on sponsored post effectiveness. Int. J. Adver.,
Evans, N. J., Phua, J., Lim, J., and Jun, H. (2017). Disclos- 41(1), 30–53. [Link]
ing Instagram influencer advertising: The effects of 1.1986257.
disclosure language on advertising recognition, atti- Leite, F. P. and de Paula Baptista, P. (2022). Influencers’ inti-
tudes, and behavioral intent. J. Interac. Adver., 17(2), mate self-disclosure and its impact on consumers’ self-
138–149. [Link] brand connections: Scale development, validation, and
66885. application. J. Res. Interac. Market., 16(3), 420–437.
Farivar, Samira, and Fang Wang. (2022). Effective influenc- [Link]
er marketing: A social identity perspective. Journal of XML.
Retailing and Consumer Services, 67: 103026, https:// Singh, S., Singh, J., and Sehra, S. S. (2020). Genetic-inspired
[Link]/10.1016/[Link].2022.103026. map matching algorithm for real-time GPS trajecto-
Farivar, Samira, Fang Wang, and Ofir Turel. (2022). Fol- ries. Arabian J. Sci. Engg., 45(4), 2587–2603.
lowers' problematic engagement with influencers on Lou, C. and Yuan, S. (2019). Influencer marketing: How
social media: An attachment theory perspective. Com- message value and credibility affect consumer trust of
puters in Human Behavior, 133: 107288, [Link] branded content on social media. J. Interac. Adver.,
org/10.1016/[Link].2022.107288. 19(1), 58–73. [Link]
Farivar, Samira, Fang Wang, and Yufei Yuan. (2021). Opin- 8.1533501.
ion leadership vs. para-social relationship: Key factors Mathur, M. B., Reichling, D. B., Lunardini, F., Geminiani,
in influencer marketing. Journal of Retailing and Con- A., Antonietti, A., Ruijten, P. A. M., Levitan, C. A.,
sumer Services, 59: 102371, [Link] et al. (2020). Uncanny but not confusing: Multisite
jretconser.2020.102371. study of perceptual category confusion in the uncanny
Haenlein, M., Anadol, E., Farnsworth, T., Hugo, H., Hu- valley. Comp. Hum. Behav., 103, 21–30. [Link]
nichen, J., and Welte, D. (2020). Navigating the new era org/10.1016/[Link].2019.08.029.
of influencer marketing: How to be successful on Ins- Mori, M., MacDorman, K. F., and Kageki, N. (2012). The
tagram, TikTok, & Co. California Manag. Rev., 63(1), uncanny valley. IEEE Robot. Autom. Mag., 19(2), 98–
5–25. [Link] 100. [Link]
Jacobson, J. and Harrison, B. (2022). Sustainable fashion Moriuchi, E. (2019). Okay, Google!: An empirical study
social media influencers and content creation calibra- on voice assistants on consumer engagement and loy-
tion. Int. J. Adver., 41(1), 150–177. [Link] alty. Psychol. Market., 36(5), 489–501. [Link]
1080/02650487.2021.2000125. org/10.1002/MAR.21192.
Jhawar, A., Kumar, P., and Varshney, S. (2023). The emer- Moustakas, Evangelos, Nishtha Lamba, Dina Mahmoud,
gence of virtual influencers: A shift in the influencer and C. Ranganathan. (2020). Blurring lines between
marketing paradigm. Young Cons., 24(4), 468–484. fiction and reality: Perspectives of experts on market-
[Link] ing effectiveness of virtual influencers. In 2020 Inter-
Karras, T., Laine, S., Aittala, M., Hellsten, J., Lehtinen, J., national Conference on Cyber Security and Protection
and Aila, T. (2019). Analyzing and improving the of Digital Services (Cyber Security), 1–6. IEEE.
image quality of StyleGAN. Proc. IEEE Comp. Soc. Nightingale, Sophie J., and Hany Farid. (2022). AI-synthe-
Conf. Comp. Vis. Patt. Recogn., 8107–8116. https:// sized faces are indistinguishable from real faces and
[Link]/10.1109/CVPR42600.2020.00813. more trustworthy. Proceedings of the National Acad-
Kim, D. Y. and Kim, H. Y. (2021). Influencer advertis- emy of Sciences. 119(8): e2120481119, [Link]
ing on social media: The multiple inference model org/10.1073/pnas.2120481119.
on influencer-product congruence and sponsorship Nissen, A. and Jahn, K. (2021). Between anthropomor-
disclosure. J. Busin. Res., 130, 405–415. [Link] phism, trust, and the uncanny valley: A dual-process-
org/10.1016/[Link].2020.02.020. ing perspective on perceived trustworthiness and its
Kim, M., Song, D., and Jang, A. (2021). Consumer response mediating effects on use intentions of social robots.
toward native advertising on social media: The roles Proc. Ann. Hawaii Int. Conf. Sys. Sci., 360–369.
of source type and content type. Internet Res., 31(5), [Link]
1656–1676. [Link] O’Brien, R. M. (2007). A caution regarding rules of thumb
0328. for variance inflation factors. Qual. Quant., 41(5),
Kolo, Castulus, and Florian Haumer. (2018). Social media 673–690. [Link]
celebrities as influencers in brand communication: An 6/METRICS.
empirical study on influencer content, its advertising Ozdemir, O., Kolfal, B., Messinger, P. R., and Rizvi, S.
relevance and audience expectations. Journal of Digi- (2023). Human or virtual: How influencer type shapes
tal & Social Media Marketing. 6(3): 273–282. brand attitudes. Comp. Hum. Behav., 145, 107771.
Lee, H. S., Sun, P. C., Chen, T. S., and Jhu, Y. J. (2015). [Link]
The effects of avatar on trust and purchase intention Park, Gyeongbin, Dongyan Nan, Eunil Park, Ki Joon Kim,
of female online consumer: consumer knowledge as Jinyoung Han, and Angel P. Del Pobil. (2021). Com-
a moderator. Int. J. Elec. Comm. Stud., 6(1), 99–118. puters as social actors? Examining how users perceive
[Link] and interact with virtual influencers on social media.
466 Statistical analysis of consumer attitudes towards virtual influencers in the metaverse
In 2021 15th International Conference on Ubiquitous Stuart, J., Aul, K., Bumbach, M. D., Stephen, A., Gomes De
Information Management and Communication (IM- Siqueira, A., and Lok, B. (2022). The effect of virtual
COM), 1–6. IEEE. humans making verbal communication mistakes on
Thakur, D., Singh, J., Dhiman, G., Shabaz, M., and Gera, learners’ perspectives of their credibility, reliability,
T. (2021). Identifying major research areas and minor and trustworthiness. 2022 IEEE Conf. Virt. Realit. 3D
research themes of android malware analysis and de- User Interf. (VR), 455–463. [Link]
tection field using LSA. Complexity, 2021, 1–28. VR51125.2022.00065.
Porra, Jaana, Mary Lacity, and Michael S Parks. (2020). Suwajanakorn, Supasorn, Steven M. Seitz, and Ira Kemelm-
Towards an Ontology and Ethics of Virtual Influ- acher-Shlizerman. (2017). Synthesizing obama: learn-
encers. Australasian Journal of Information Systems, ing lip sync from audio. ACM Transactions on Graph-
24 (June), 1–8. [Link] ics (ToG), 36(4): 1–13.
2807. Tan, S. M. and Liew, T. W. (2020). Designing embodied vir-
Rutter, R. N., Barnes, S. J., Roper, S., Nadeau, J., and Lettice, tual agents as product specialists in a multi-product
F. (2021). Social media influencers, product placement category e-commerce: The roles of source credibility
and network engagement: Using AI image analysis and social presence. Int. J. Human-Comp. Interac.,
to empirically test relationships. Indus. Manag. Data 36(12), 1136–1149. [Link]
Sys., 121(12), 2387–2410. [Link] 8.2020.1722399.
IMDS-02-2021-0093. Uyanık, G. K. and Güler, N. (2013). A study on mul-
Sands, S., Campbell, C. L., Plangger, K., and Ferraro, C. tiple linear regression analysis. Proc. Soc. Behav.
(2022). Unreal influence: Leveraging AI in influencer Sci., 106, 234–240. [Link]
marketing. Eur. J. Market., 56(6), 1721–1747. https:// SPRO.2013.12.027.
[Link]/10.1108/EJM-12-2019-0949. Yew, Roy Ling Hang, Syamimi Binti Suhaidi, Prishtee See-
Oliveira, S., Batista da, A., and Chimenti, P. (2021). ‘Human- woochurn, and Venantius Kumar Sevamalai. (2018).
ized Robots’: A proposition of categories to under- Social network influencers’ engagement rate algorithm
stand virtual influencers. Australasian J. Inform. Sys., using instagram data. In 2018 fourth international
25, 1–27. [Link] conference on advances in computing, communication
Sokolova, Karina, and Hajer Kefi. (2020). Instagram and & automation (icacca), 1–8. IEEE.
YouTube bloggers promote it, why should I buy? Zhao, Y., Jiang, J., Chen, Y., Liu, R., Yang, Y., Xue, X.,
How credibility and parasocial interaction influence and Chen, S. (2022). Metaverse: Perspectives from
purchase intentions. Journal of retailing and consumer graphics, interactions and visualization. Visual In-
services, 53: 101742, [Link] format., 6(1), 56–67. [Link]
conser.2019.01.011. VISINF.2022.03.002.
60 Quantum dynamics-aided learning for secure integration
of body area networks within the metaverse cybersecurity
framework
Anand Singh Rajawat1, S. B. Goyal2,a, Jaiteg Singh3 and Celestine Iwendi4
1
School of Computer Science & Engineering, Sandip University, Nashik, Maharashtra, India
2
City University, Petaling Jaya, 46100, Malaysia
3
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
4
SMIEEE, School of Creative Technologies, University of Bolton, United Kingdom
Abstract
The seamless integration of body area networks (BAN) poses several cybersecurity challenges within the continuously de-
veloping metaverse. This research proposes a novel technique that combines quantum dynamics, especially quantum meta-
verse (QMV), with the conventional dynamics of the BAN system (SBAN). The aim is to enhance security and facilitate the
learning process. In order to guarantee the secure integration of personal and biometric data acquired via BANs, the authors
propose the utilization of a quantum dynamics-aided learning framework. This model serves as a connection between the
realm of swiftly advancing quantum computing and the growing demands of the metaverse. The enhancement of intrusion
detection capabilities is just one aspect of our methodology that demonstrates its effectiveness in mitigating the dangers as-
sociated with integrating BAN data inside intricate virtual environments. The effectiveness of the model in mitigating diverse
cyber threats has been demonstrated through rigorous evaluation in simulated environments as well as real-world scenarios.
The findings indicate an initial stride towards establishing a safer and more immersive setting for metaverse users, while also
addressing the pressing demand for enhanced cybersecurity protocols.
Keywords: Quantum dynamics, body area networks, metaverse cybersecurity, secure integration, quantum metaverse, secu-
rity for body area networks
a
drsbgoyal@[Link]
468 Quantum dynamics-aided learning for secure integration of body area networks
T. Nussle and J. Path integral method Potentially more accurate Not specified based Depth and breadth of
Barker, 2023 for simulations of spin simulation of spin dynamics on provided info simulation scenarios
dynamics
M. Ballicchia, M. Wigner dynamics for Enhanced understanding of Limitation in Extent of
Nedjalkov and J. electron quantum quantum states in confined scalability might applicability to other
Weinbub, 2022 superposition states and opened quantum dots exist systems
T. Itami, N. Quantum computation New approach to quantum Likely less efficient Full realization
Matsui and T. by classical mechanical computation using classical than pure quantum and potential
Isokawa, 2020 apparatuses mechanics methods optimizations
S. Chen, 2022 Quantum computer Unified analysis of cultivated Possible high Integration with
assisted dynamics & ecological land with computational costs broader ecological
modeling quantum computer support models
Applied Data Science and Smart Systems 469
• Maximize S(BAN)
• Subject to Q(MV)
within the context of integrating BANs (Tonmoy et Q(t) -> Controller -> S(t) -> BAN -> Q(t+1)
al., 2020) within the cyberspace of the metaverse. The
utilization of quantum dynamics-assisted learning has The controller receives the quantum dynamics Q(t) as
the potential to enhance security measures in BANs its input and produces the freshly configured security
(Stitely et al., 2022), therefore safeguarding the pri- state S(t) of the BAN as its output. The altered secu-
vacy of users’ personal data within the metaverse. rity configuration has a subsequent influence on the
The proposed study aims to develop a model for evolution of the system, leading to a distinct set of
the secure integration of BANs (Dong et al., 2021) quantum dynamics denoted as Q(t+1).
within the metaverse cybersecurity framework. This In order to maintain the ongoing security of the
model considers the BAN’s security and the quantum blockchain autonomous network (Langenickel et
dynamics of the metaverse environment as influential al., 2021) within the metaverse, the learning process
factors in a dynamic process that evolves over time incorporates the principles of quantum dynamics,
(Morishita et al., 2023). The study proposes the use of denoted as S(BAN)+ Q(MV). One possible approach
quantum dynamics (S(BAN) + Q(MV))-aided learn- to tackle this issue is by formulating it as an optimiza-
ing to represent and analyze this system. tion problem:
The quantum dynamics of the metaverse environ-
ment will be denoted as Q(t), while the security of the • maximize ∫[0,∞] S(t) dt
BAN will be denoted as S(t). A differential equation • subject to Q(t).
has the ability to depict the temporal evolution of a
system’s behavior (Jin et al., 2022):
The optimal approach to managing the security con-
figuration of the body area network in light of the
dS/dt = f(S(t), Q(t)) observed quantum dynamics inside the metaverse
environment can be determined through the resolution
where f is a function that captures the interplay of the associated optimization issue. The safeguarding
between security and quantum dynamics. This func- of users’ personal information will be ensured within
tion can be further decomposed into two components: the metaverse (Cranganore et al., 2022).
The task of constructing a comprehensive table that
f(S(t), Q(t)) = g(S(t)) + h(Q(t)) outlines the datasets used in the research on “Quantum
Dynamics (S(BAN) + Q(MV))-Aided Learning for
The variable “g” represents the security dynamics Secure Integration of Body Area Networks within the
that are inherent to the BAN, while the variable “h” Metaverse Cybersecurity Framework” is a complex
represents the influence of the quantum metaverse on and highly specialized endeavor (Sarantoglou et al.,
the aforementioned security. 2020). As of January 2022, there is a lack of train-
The notion of quantum dynamics (S(BAN) + ing data available that precisely aligns with the speci-
Q(MV))-aided learning can be seen as an adaptive fied standards, likely due to the novelty of the subject
control mechanism that adjusts the security configu- matter. Individuals have the capacity to independently
ration of the BAN in accordance with variations in gather pertinent datasets.
the quantum dynamics of the metaverse. A feedback Table 60.2 functions as a schematic (Angelopoulou,
loop might be employed to show this phenomenon. 2023). To facilitate the execution of your research, it
BAN security dataset E.g., Time-series, Data related to threats and To study common threats and
50 GB vulnerabilities in body area networks devise quantum-aided solutions
Metaverse E.g., Graph data, Data representing user interactions For understanding patterns and
interaction 100 GB within a metaverse potential vulnerabilities
Quantum dynamics E.g., Logs, 25 GB Logs from quantum devices or Aiding the learning model
logs simulations showing quantum in understanding quantum
dynamics behaviors
Cybersecurity E.g., Relational DB, Established cybersecurity practices and For integrating BAN securely
frameworks dataset 10 GB protocols for different networks within the metaverse framework
User behavioral data E.g., Time-series, Data depicting user behavior in both To identify potential misuse or
75 GB BAN and metaverse environments anomalous behaviors
Applied Data Science and Smart Systems 471
Table 60.3 Simulation Parameters for Assisted Learning in Quantum Dynamics (S(BAN)+ Q(MV))
Quantum parameters
Qubit number Number of quantum bits used 5
Quantum gate set The set of quantum gates used {X, Y, Z, H}
Quantum circuit depth Number of operations in the quantum circuit 20
Noise model Model for quantum noise Depolarizing
Decoherence time Time before qubits lose coherence 10 μs
Body area network (BAN) parameters
Node number Number of nodes in the BAN 10
Transmission power Power for data transmission -10 dBm
Data rate Rate of data transmission 250 kbps
Sensing frequency Frequency of data collection 10 Hz
472 Quantum dynamics-aided learning for secure integration of body area networks
employ the previously described simulation settings Table 60.4 presented herein exhibits fabricated
and present their outcomes or performance metrics in data derived from the parameters outlined in
a tabular format during the analysis of results. Based the preceding inquiry. Empirical simulations and
on the aforementioned inputs, I thus present the next evaluations are necessary to ascertain real-world
illustration (Yong, 2021). performance.
Table 60.4 Analytical results on the utilization of quantum dynamics to enhance learning in a multi-agent environment
Quantum parameter
Qubit number 5 95% accuracy in computation Satisfactory performance
Quantum gate set {X, Y, Z, H} Minimal gate errors (<0.01%) Robust gate set
Quantum circuit depth 20 Average 0.05% error rate Stable for this depth
Noise model Depolarizing Affects 1 in every 100 computations Need error correction
Decoherence time 10μs No loss in 98% of operations Optimal performance
BAN parameters
Node number 10 98% successful data collection Good node connectivity
Transmission power -10 dBm 95% packets received without distortion Adequate power
Data rate 250 kbps Minimal data congestion (3%) Efficient rate
Sensing frequency 10 Hz 99% uptime in sensing Consistent sensing
Metaverse cybersecurity framework
Attack model DDoS 90% attacks mitigated Further fortification needed
Encryption algorithm AES-256 100% secure data transmissions Highly secure
Key exchange protocol ECDH 99.9% secure key exchanges Reliable exchange
Anomaly detection ML-based 95% anomalies detected Effective detection
mechanism
Integration parameters
Integration latency 5ms e.g., 98% successful real-time integrations Minimal delays
Data transfer rate 1 Gbps e.g., 97% bandwidth utilization Optimal transfer
Synchronization NTP e.g., 99.5% synced operations Almost perfect sync
mechanism
Learning rate 0.001 e.g., Convergence after 450 epochs Appropriate learning
Epochs 500 e.g., 96% learning efficiency Good training duration
Applied Data Science and Smart Systems 473
Abstract
The need for wireless networks among users and their unique features make FANETs an attractive and emerging technology.
FANET-based research and development, both in academia and business, has surged in recent years. Due to their unique
qualities for many vital mission applications, unmanned aerial vehicles (UAVs) are being used more and more for a range
of missions, such as traffic surveillance, video graphics, and military, and civilian operations. The research proposes a back
propagation neural network technique based on supervised learning for clustering-based location-aware and energy-efficient
routing. The suggested results show that, for FANET, the network lifetime effectively rises and the energy consumption is
somewhat decreased. A developing method for estimating network performance and achieving an energy-efficient solution
which is achieved by utilizing the suggested approach is supervised learning.
jindal08@[Link]
476 An optimized approach for development of location-aware-based energy-efficient routing for FANETs
change. As a result, deploying UAVs as an interme- monitoring (Albu-Salih et al., 2021; Da Silva et al.,
diary node in already-existing ad-hoc networks can 2021). Therefore, it becomes crucial to join several
effectively handle challenging jobs. Cooperative UAVs to create an independent aerial network that
search, object tracking, data collecting, and data can work in tandem with current ground networks
analysis are some of these challenging activities. (Figure 61.2).
UAVs can be deployed either singly or in groups. A
single UAV system has been effectively coordinated Motivation
with pre-existing ad-hoc formations in the literature.
However, single UAV coordinated networks struggle Flying ad-hoc networks (FANETs) or unmanned aer-
with scalability and can only offer a modest level of ial vehicular networks are currently facing additional
issues in terms of energy efficiency, as a result of the
explosive development in traffic demand from users
for a variety of services such as traffic surveillance,
live video streaming, health care monitoring pur-
poses, etc. Additionally, because all of these issues are
entirely dependent on the routing method in FANETs,
they grow more serious as more secure and energy-
efficient services, such as traffic surveillance and mili-
tary applications, are demanded. So energy efficiency
is a critical part of FANETs. As a result, routing in
FANETs has recently attracted a lot of attention from
the research community. However, to solve FANETs’
problems using relevance clustering along with the
idea of hybrid optimization technique with artificial
intelligence technique, researchers also need to take
into account other factors like random deployment,
network security problems, and energy efficiency.
To overcome the current challenging factor of the
FANETs and improve the quality of service (QoS) in
Figure 61.1 Different UAV network level communica- terms of throughput, end-to-end delay, packet delivery
tions ratio, packet loss rate, collision avoidance intensity,
energy consumption, location awareness, and energy- adaptive epsilon-greedy strategy. Additionally, the
efficient routing mechanism based on optimized artifi- modeling findings demonstrate its usefulness having
cial intelligence (AI) technique was chosen (Bhardwaj a 39.9% greater detection rate than other methods
and Kaur, 2021; Al-Absi et al., 2021). or algorithms in the period of FANETs. A mobility-
assisted adaptive routing for FANETs made up of
Related work several UAVs that are intermittently connected was
presented by Li et al. (2020) and Singh et al. (2021).
Ali et al. (2021) researched an architecture designed Because the current routing algorithms are insuffi-
for routing in flying ad-hoc networks (FANETs), cient for mobility-based networks, the authors of this
which is crucial for optimizing the use of drone- study introduced mobility assisted adaptive routing
based internet in various everyday applications. (MAAR), a geographic routing method. Unlike tra-
These applications range from monitoring traffic ditional routing protocols in FANETs that rely on
and agriculture to aiding in healthcare, managing location services for gathering location information,
disasters, and assisting in various rescue missions. the MAAR algorithm integrates a routing strategy
Nonetheless, the dynamic nature and constant topo- with a location service.
logical changes in UAVs present significant challenges This approach aims to decrease both the latency
in FANETs, particularly in selecting the appropriate and the overhead involved in routing data packets.
next node, adapting autonomously, and preventing They adopt the store-carry-and-forward paradigm
the formation of routing loops. The performance to address the technological problems posed by net-
of a FANET should be significantly improved for works that experience communication outages. For
future implementation. As a result, the authors of FANETs, which are time-varying networks with
this study created a performance-aware routing sys- dynamic links that make it challenging to sustain
tem for effective UAV-to-UAV communication in a constant communication. Sang et al. (2020) presented
FANET context. Liu (2019) conducted research an energy-efficient opportunistic routing strategy in
on the FANETs’ performance-aware routing archi- 2020. The EORB-TP protocol, which was proposed
tecture. It’s a technique for realizing the potential by the authors, is a new trajectory prediction-based
of the Internet of Drones in a variety of everyday opportunistic routing system. The idea of resource-
applications, such as traffic surveillance, agricultural ful communication was utilized to resolve the issue of
monitoring, the healthcare system, disaster manage- different uncertainties that depends on the node archi-
ment, and countless rescue operations. However, due tecture, which allowed for the prediction of the posi-
to UAVs rapid movements and frequent topological tion of UAV. To prevent overconsumption, the node’s
modifications, choosing the next hop, allowing for trajectory metric value was then calculated based
self-adaptation, and avoiding dissemination loops on the UAVs or the node’s trajectory parameters.
have proven to be difficult problems in FANETs. For As wireless connectivity was a significant problem
use in the future, a FANET’s performance needs to be in a particular coverage region Tropea et al. (2020)
greatly enhanced. did research on the FANET simulator for managing
To facilitate efficient UAV-to-UAV communication drones and enabling dynamic connectivity in the net-
in a FANET environment, the authors of this paper work. The authors of this study attempt to deal with
developed a performance-aware routing system. Due these new types of flying ad-hoc networks that might
to the wireless nature of FANETs and the particu- be appropriate for any emergencies where the classic
lar network 16 features (Mowla et al., 2020) devel- networking paradigm may encounter several prob-
oped an adaptive federated reinforcement learning lems or implementation challenges. With the develop-
(AFRL) mechanism for intelligent jamming defense. ment of a UAV/drone behavior model to account for
Before taking into account the mobility density of drones’ energetic concerns, the goal of this work was
the UAVs, the authors first made a decision based to build new methods of area coverage and human
on a centralized knowledge base on the commu- movement behaviors.
nication and power limits in FANET. Finally, in a
recently investigated environment, a model-based
jamming defense action was constructed and an
Problem statement
AFRL-based jamming attack defense plan was pro- In networks of unmanned aerial vehicles, com-
vided. An innovative jamming detection system for monly known as UAVNs or flying ad-hoc networks
flying ad-hoc networks (FANETs) has been devel- (FANETs), there is a growing concern about energy
oped using a Q-learning approach that doesn’t rely efficiency and network longevity. This is primarily
on pre-existing models. This system enhances its due to the unpredictable positioning and diminishing
performance by dynamically adjusting the balance range of a large number of drones, which negatively
between exploration and exploitation, utilizing an impacts the quality of communication. Due to their
478 An optimized approach for development of location-aware-based energy-efficient routing for FANETs
Department of Computer Science, Lord Buddha Education Foundation-LBEF Campus, Kathmandu, Nepal
3
Abstract
Despite ongoing advancements, certain complex computational tasks still face challenges in scalability and performance
within the existing cloud computing paradigm. This study investigates the integration of Quantum Monte Carlo (QMC) and
quantum machine learning (QML) methodologies into cloud architectures. Quantum Monte Carlo, a probabilistic method-
ology, leverages quantum principles to effectively and precisely address intricate systems. Quantum machine learning (QML)
leverages principles from quantum physics to enhance the computational efficiency of machine learning algorithms, leading
to substantial reductions in processing time and enhanced predictive accuracy. By integrating these quantum algorithms into
cloud systems, we are able to demonstrate enhanced scalability and resilient performance, even when subjected to substantial
workloads. In order to address the existing limitations of conventional cloud systems and pave the path for future advance-
ments in the integration of quantum computing with cloud technologies, a framework known as quantum cloud computing
was proposed. Initial trials demonstrate potential, instilling optimism that quantum cloud computing could provide a novel
epoch of expeditious digital metamorphosis and enhanced computational capacities spanning many domains.
Keywords: Quantum parallelism, quantum entanglement, quantum superposition, quantum Monte Carlo simulations, quan-
tum neural networks (QNNs), quantum cloud infrastructure
drsbgoyal@[Link]
a
Applied Data Science and Smart Systems 483
Figure 62.1 Integrating quantum model for enhanced scalability and performance in cloud architectures
Quantum computing (entanglement, superposition, The quantum computation’s intermediate or raw re-
etc.) – Use entanglement and superposition to sults should be sent to the classical cloud infra-
solve problems. structure.
Applied Data Science and Smart Systems 485
Classical post-processing and data synthesis – Clas- Collect data on system performance and user
sical systems examine, synthesize, and perhaps feedback.
process quantum data. Refine quantum algorithms and integration layers.
Quantum computing improves cloud service delivery Scale quantum resources based on demand.
– Users receive the finished service or data.
Feedback loop
Description of how the flowchart might look are
as follows: The flowchart should include a feedback loop from
deployment and continuous improvement stages
Identify quantum-ready processes back to the design and assessment stages for iterative
Start with identifying processes that would bene- enhancements.
fit from quantum computing (Zhao et al., 2020). The visualization of step-by-step flowchart process
Determine scalability and performance require- is given below (Zimmermann et al., 2013): End
ments. User receives enhanced services and feedback loop.
Evaluate the compatibility of current cloud ar- The user benefits from the enhanced services, and
chitecture with quantum processes. their feedback or further requests may be fed back
Assess quantum computing resources into the system for continuous improvement.
Identify available quantum computers or quan- Identify the problem or application that can be ben-
tum cloud services (like IBM Q, Rigetti, etc.). efited by using QMC and QML algorithms (O’Meara
Determine quantum processing power (Qubits, et al., 2023).
quantum volume). QMC and QML algorithms are highly suitable for
Assess quantum programming languages addressing challenges such as describing the behav-
(QASM, Qiskit, etc.). ior of intricate molecules and materials, developing
Design quantum algorithms innovative machine learning models, and address-
Translate identified processes into quantum al- ing intricate optimization problems. The appropriate
gorithms. QMC and QML algorithms is selected (Chekired et
Optimize algorithms for the specific quantum al., 2017).
processor. There is a wide range of quantum Monte Carlo
Use quantum simulation tools for testing. (QMC) and QML methods available, each possessing
Integrate quantum algorithms with cloud services distinct merits and drawbacks. The thorough selec-
Develop APIs for integration of quantum algo- tion of algorithms is crucial in order to assure their
rithms with existing cloud services (Pourvahab optimality for the given task (Figure 62.2).
and Ekbatanifard, 2019). The QMC and QML algorithms are developed or
Ensure data security during quantum processing. adopted to run on a quantum computer.
Set up a hybrid cloud-quantum environment. Most quantum Monte Carlo (QMC) and QML
Scalability planning methods are primarily designed and optimized for
Design systems for easy scaling of quantum re- implementation on classical computing systems. The
sources. quantum computer may require certain adjustments
Implement a microservices architecture to encap- in order to ensure optimal functionality of the given
sulate quantum processes. components.
Plan for quantum error correction and fault tol- The QMC and QML algorithms are integrated
erance. with a cloud computing platform.
Performance benchmarking Cloud computing systems provide (Ramidi et al.,
Compare quantum-enhanced processes with 2017) the accessibility of quantum computers and
classical processes. facilitate the scalability of quantum Monte Carlo
Record time and resource efficiency improve- (QMC) and QML algorithms to effectively handle
ments. substantial workloads.
Adjust quantum algorithms based on perfor- The QMC and QML algorithms are deployed and
mance data. run on the cloud computing platform.
Deploy quantum-enhanced cloud services The successful delivery and execution of the prob-
Roll out quantum-enhanced services to end-us- lem or application is contingent upon the integration
ers. of QMC and QML with the cloud computing plat-
Monitor system performance and stability. form (Barcelo et al., 2016).
Provide support for quantum-based applica- By employing this methodology, the quantum Monte
tions. Carlo (QMC) algorithm may be seamlessly included
Continuous improvement and scaling into quantum cloud computing infrastructure, hence
486 Quantum cloud computing: Integrating quantum algorithms for enhanced scalability
facilitating the acceleration of the drug development Develop or adapt the QMC algorithms to run on a
process. quantum computer.
QMC algorithms are commonly designed with a
Identify the problem or application that can be ben- focus on classical computing systems. The quan-
efited by using QMC algorithms. tum computer may require several adjustments
QMC algorithms have the capability to simulate in order to ensure appropriate functionality.
complex molecules, including medicinal com- Integrate the QMC algorithms with a cloud comput-
pounds. Both the advancement of existing medi- ing platform.
cations and the creation of new pharmaceuticals Quantum Monte Carlo (QMC) algorithms pos-
can derive advantages from the above. sess the capability to be expanded in order to han-
Choose the appropriate QMC algorithms. dle substantial workloads and can be convenient-
Various quantum Monte Carlo (QMC) algo- ly accessed through cloud computing platforms.
rithms have distinct strengths and weaknesses. Deploy and run the QMC algorithms on the cloud
The selection of an efficient and extensible ap- computing platform.
proach is of utmost importance in the drug de-
velopment process. Before the deployment and operation of QMC
algorithms for modeling the behavior of medicinal
Applied Data Science and Smart Systems 487
By concurrently executing the quantum Monte Quantum cloud computing, an emerging field
Carlo (QMC) and QML (Dudhe et al., 2018) algo- that integrates quantum Monte Carlo (QMC) and
rithms on many quantum computers, it becomes fea- QML techniques, is currently in its nascent stage
sible to simulate larger and more intricate systems. but holds significant promise to revolutionize vari-
Cloud computing systems enable the integration of ous industries. The acceleration of pharmaceutical
QML and quantum computation (QMC) techniques and material innovation can be facilitated through
across several quantum computers. the utilization of quantum Monte Carlo (QMC)
Cloud infrastructures can provide the necessary algorithms. Similarly, QML algorithms offer the
computational resources and accommodate the large potential for the development of novel applications
datasets required for training and deploying QMC in machine learning. Furthermore, the applica-
and QML models. tion of both QMC and QML algorithms presents
The utilization of cloud architectures can prove an opportunity to address complex challenges in
advantageous for QMC and QML algorithms as it various sectors such as logistics, finance, and other
facilitates the distribution of computational load industries.
across multiple quantum computers, hence granting An illustration of the formula’s application is pre-
users access to substantial computing capabilities. sented below:
Table 62.2 A comprehensive overview of the impact of each variable on the main performance metrics
Parameter Varied values Impact on quantum Impact on fidelity Impact on Impact on training
name speedup quantum volume convergence
Quantum bits 10, 20, 30, 40 Increase with more Decrease due Increase with Slower convergence
(qubits) qubits to increased more qubits, but with more qubits
complexity plateaus
Noise level Low, medium, Decrease with Significant Decrease with Slower convergence
high higher noise decrease with higher noise and possible non-
higher noise convergence with
high noise
Decoherence 50, 100, 200 Increase with Increase Slightly improved Faster convergence
time microseconds longer decoherence with longer with longer with longer
time decoherence time decoherence time decoherence time
Gate fidelity 0.95, 0.97, 0.99 Increase with Significant Increase with Faster convergence
higher fidelity increase with higher fidelity with higher fidelity
higher fidelity
Trotter steps 200, 500, 800 Marginal speedup Improved fidelity Little to no Convergence
(QMC) with more steps with more steps impact improves with more
steps but plateaus
Sampling rate 20, 50, 80 Increased speedup Slight Little to no Faster convergence
(QMC) samples/step with more samples improvement in impact with more samples
fidelity with more
samples
Walkers 200, 500, 800 Improved speedup Increased fidelity Marginal impact Improved
(QMC) with more walkers with more convergence rate
walkers with more walkers
Training data 100, 500, 900 Speedup plateaus Fidelity increases Little to no direct Faster convergence
size (QML) after a certain size with more data, impact initially, but
but plateaus marginal gains after
a threshold
Quantum 2, 5, 8 Speedup increases Fidelity improves Marginal impact Slower convergence
layers (QML) with more layers but then starts with more layers
but plateaus declining due
to increased
complexity
Parameterized RX, RY vs. RX, Better speedup with Improved fidelity No direct impact Slight delay in
gates (QML) RY, CNOT more varied gates with diverse gates convergence with
more complex gates
Applied Data Science and Smart Systems 489
Currently, there is ongoing development of a drug predictive capabilities gives rise to a powerful com-
discovery algorithm that utilizes the unique approach putational framework capable of handling intricate
of quantum Monte Carlo (QMC). The firm does not quantum states, efficiently analyzing extensive quan-
possess intentions to engage in the development or tum datasets, and enhancing optimization in many
upkeep of its own quantum computers; nonetheless, it application domains.
does aspire to facilitate widespread access to the algo- Cloud architectures possess the capability to
rithm within academic circles. To integrate its quan- address a diverse range of difficulties, spanning
tum Monte Carlo (QMC) algorithm into the cloud from quantum chemistry to optimization, owing
provider’s infrastructure, the company has opted to to their unified approach in harnessing quantum
establish a collaborative alliance with the aforemen- resources. Quantum cloud computing (QMC) is
tioned provider. The connection would enable the poised to initiate a paradigm shift in high-perfor-
company to globally distribute its QMC algorithm mance computing, surmounting numerous limita-
to consumers and effectively expand its capacity to tions inherent in classical cloud infrastructures
accommodate a substantial user population. through the synergistic use of QMC’s precision and
By implementing this interface, the organiza- QML’s adaptability.
tion would be able to leverage the machine learning However, there are other challenges that must be
capabilities of the cloud service provider to develop overcome in order to fully realize the potential of
advanced QML algorithms. The company might quantum cloud computing. Significant efforts are
potentially leverage the QML algorithm development still required to address the challenges pertaining to
capabilities of the cloud service provider to facilitate quantum noise, decoherence, and the dependability
the training and deployment of novel drug discovery of quantum gates. Nevertheless, advancements in
models. quantum error correction and mitigation techniques
The integration of quantum cloud computing with offer a basis for optimism that these challenges can
QMC and QML algorithms holds significant poten- be surmounted.
tial for enhancing the drug discovery industry. The integration of quantum Monte Carlo (QMC)
and QML in the field of quantum cloud computing
Results analysis is gaining attention as a potential solution to address
the growing need for enhanced computational capa-
To conduct an analysis of the results obtained from a bility, as well as the expanding boundaries of conven-
simulation run utilizing the previous parameter table, tional computing. Despite being in its early stages,
it is necessary to build a new table (Table 62.2). quantum cloud computing exhibits significant poten-
The results presented in Table 62.2 are hypotheti- tial for transforming various domains of research and
cal and should not be interpreted as representative commerce. The integration of QMC (quantum Monte
of actual outcomes. Based on the aforementioned Carlo) with QML signifies a significant paradigm shift
dimensions, it is evident that alterations in any of in comprehending and addressing the vast capabilities
these factors can potentially impact the remaining of cloud computing.
performance indicators. The final findings can be
significantly influenced by various factors, including
the specifics of the simulation, the quantum technol- References
ogy employed, and the nature of the activity being Li, C., Guo, Z., He, X., Hu, F., and Meng, W. (2023). An
undertaken. AI model automatic training and deployment plat-
form based on cloud edge architecture for DC energy-
Conclusion saving. 2023 Int. Conf. Mob. Internet Cloud Comput.
Inform. Sec. (MICCIS), 22–28.
The integration of quantum Monte Carlo (QMC) and Volpe, G., Mangini, A. M., and Fanti, M. P. (2022). An ar-
QML methodologies into cloud architectures signi- chitecture combining blockchain, docker and cloud
fies a significant advancement in the progression of storage for improving digital processes in cloud manu-
cloud infrastructures. The combination of inherent facturing. IEEE Acc., 10, 79141–79151.
quantum parallelism and the quantum-mechanical Jiang, F.-C., Hsu, C.-H., and Wang, S. (2016). Logistic sup-
port architecture with petri net design in cloud envi-
properties of qubits has great potential for achieving
ronment for services and profit optimization. IEEE
significant scalability and performance advantages. Trans. Ser. Comput., 10(6), 879–888.
The stochastic simulation of quantum states by El Mhouti, A., Mohamed Erradi, A. N., and Vasquèz, J. M.
quantum Monte Carlo (QMC) offers significant (2016). Cloud-based VCLE: A virtual collaborative
insights, particularly in situations when classical sys- learning environment based on a cloud computing ar-
tems encounter difficulties in accurately simulating chitecture. 2016 Third Int. Conf. Sys. Collab. (SysCo),
such states. The incorporation of QML’s adaptive and 1–6.
490 Quantum cloud computing: Integrating quantum algorithms for enhanced scalability
Radzid, A. R., Azmi, M. S., Jalil, I. E. A., Mas’ ud, M. Z., Ramidi, D. R., Katangur, A. K., and Kar, D. C. (2017). Vir-
Arbain, N. A., and Melhem, L. B. (2018). Architecture tual machine migration and task mapping architec-
of resource management in the cloud environment: ture for energy optimization in cloud. 2017 Int. Conf.
Review and proposed of ViDaC. 2018 Int. Conf. Elec Comput. Sci. Comput. Intel. (CSCI), 1566–1571.
Con Optim. Comp. Sci. (ICECOCS), 1–6. Gera, T., Singh, J., Mehbodniya, A., Webber, J. L., Shabaz,
Wang, L. (2019). Architecture-based reliability-sensitive M., and Thakur, D. (2021). Dominant feature selec-
criticality measure for fault-tolerance cloud applica- tion and machine learning-based hybrid approach
tions. IEEE Trans. Paral. Distrib. Sys., 30(11), 2408– to analyze android ransomware. Sec. Comm. Netw.,
2421. 2021, 1–22.
Zhao, S., Wang, J., Zhang, J., Bao, J., and Zhong, R. (2020). Barcelo, M., Correa, A., Llorca, J., Tulino, A. M., Vicario, J.
Edge-cloud collaborative fabric defect detection based L., and Morell, A. (2016). IoT-cloud service optimiza-
on industrial internet architecture. 2020 IEEE 18th tion in next generation smart environments. IEEE J.
Int. Conf. Indus. Inform. (INDIN), 1, 483–487. Sel. Areas Comm., 34(12), 4077–4090.
Pourvahab, M. and Ekbatanifard, G. (2019). Digital fo- Kjamilji, A. (2014). Multi-objective optimizations during
rensics architecture for evidence collection and prov- parallel processing in a dynamic heterogeneous cloud
enance preservation in IAAS cloud environment us- environment. 2014 Sixth Int. Conf. Comput. Intel.
ing SDN and blockchain technology. IEEE Acc., 7, Comm. Sys. Netw., 131–138.
153349–153364. Jain, V. and Kumar, B. Optimal task offloading and resource
Zimmermann, A., Pretz, M., Zimmermann, G., Firesmith, allotment towards fog-cloud architecture. 2021 11th
D. G., Petrov, I., and El-Sheikh, E. (2013). Towards Int. Conf. Cloud Comput. Data Sci. Engg. (Conflu-
service-oriented enterprise architectures for big data ence), 233–238.
applications in the cloud. 2013 17th IEEE Int. Enterp. Alla, H. B., Alla, S. B., and Ezzati, A. (2016). A novel archi-
Distrib. Object Comput. Conf. Workshops, 130–135. tecture for task scheduling based on dynamic queues
O’Meara, C., Fernández-Campoamor, M., Cortiana, G., and particle swarm optimization in cloud computing.
and Bernabé-Moreno, J. (2023). Quantum software 2016 2nd Int. Conf. Cloud Comput. Technol. Appl.
architecture blueprints for the cloud: Overview and (CloudTech), 108–114.
application to peer-2-peer energy trading. 2023 IEEE Dudhe, A., Sherekar, S. S., and Thakare, V. M. Critical
Conf. Technol. Sustain. (SusTech), 191–198. analysis of performance optimization of mobile web
Chekired, D. A., Khoukhi, L., and Mouftah, H. T. (2017). services in cloud environment. 2018 3rd Int. Conf.
Decentralized cloud-SDN architecture in smart grid: Comm. Elec. Sys. (ICCES), 355–360.
A dynamic pricing model. IEEE Trans. Indus. Inform.,
14(3), 1220–1231.
63 Integrating AI-enabled post-quantum models in
quantum cyber-physical systems opportunities
and challenges
S. B. Goyal1,a, Anand Singh Rajawat2, Ruchi Mittal3 and
Divya Prakash Shrivastava4
1
School of Computer Science & Engineering, Sandip University, Nashik, Maharashtra, India
2
City University, Petaling Jaya, 46100, Malaysia
3
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
4
Department Computer Science, Higher Colleges of Technology, Dubai, United Arab Emirates
Abstract
The convergence of traditional cyber-physical systems (CPS), quantum computing, and artificial intelligence (AI) gives rise
to a novel system known as a quantum cyber-physical system (QCPS). This study aims to examine the integration of post-
quantum models enabled by AI into quantum computing platforms and systems (QCPS). The merging of AI methodologies
and the computational capabilities of quantum computers presents a novel approach to addressing intricate challenges in
CPS. This context has several potential outcomes, including enhanced safety measures, improved resource allocation, and in-
creased efficiency in quantum operations. Nevertheless, it is imperative to meticulously examine several challenges that arise
in this context, including quantum decoherence, the interpretability of AI models, and the nascent stage of post-quantum
algorithms. Overcoming these challenges will facilitate the advent of a novel era characterized by the integration of quantum-
enabled systems, hence holding the capacity to revolutionize numerous domains within the economy and societal structure.
Keywords: Quantum computing (QC), artificial intelligence (AI), post-quantum cryptography (PQ), cyber-physical systems
(CPS), integration challenges, quantum opportunities
a
drsbgoyal@[Link]
492 Integrating AI-enabled post-quantum models in quantum cyber-physical systems
of quantum computation and cryptography. In their the possible applications of this technology as well as
seminal study, Tosh et al. (2020) undertook a sig- the challenges that need to be addressed prior to its
nificant research endeavor aimed at using quantum extensive implementation.
computing techniques to enhance the security of The study conducted by Vereno et al. (2023) exam-
cyber-physical systems. The investigation of quantum ined the potential of quantum power flow algorithms
algorithms has been conducted within the frame- in enhancing energy distribution optimization within
work of safeguarding these systems against diverse the context of smart grids. The study conducted by
cyberattacks. the researchers showcased the potential of quantum
Numerous studies have been conducted to exam- algorithms in simulating and controlling energy dis-
ine the possibilities of quantum cryptography in tribution within smart grids. This discovery presents
safeguarding cyber-physical systems, with a special a promising avenue for improving the efficiency and
emphasis on smart grids. Zhang et al. (2015) exten- reliability of these critical infrastructures.
sively examined the utilization of quantum cryptog- Each article has the potential to contribute to the
raphy-based security methods specifically tailored for creation of a comprehensive table that summarizes
smart grids, emphasizing their efficacy in safeguarding its methods, advantages, limitations, and areas for
communication channels from unauthorized access further investigation. It is important to note that the
and manipulation. Consequently, these techniques comprehensiveness and accuracy of Table 1 are con-
contribute to the enhanced stability and resilience of tingent upon the data provided. Without a careful
power grids. examination of the complete articles, the table may
The authors Rajawat et al. (2022) provided a only offer a limited perspective.
detailed account of a newly developed cyber-physical Table 63.1 provides a comprehensive summary
system designed for industrial automation, which based on the titles, presumed methodologies, advan-
integrates principles from both quantum physics and tages, disadvantages, and gaps. In order to achieve
artificial intelligence. The suggested system utilizes a comprehensive understanding, it is important to
quantum deep learning algorithms to enhance the engage in a thorough examination of each object,
efficiency and safety of automation, hence enabling demonstrating attentiveness to the specific particu-
the achievement of effective manufacturing and pro- lars. It is imperative to conduct a thorough evaluation
duction systems. of each source in order to identify and implement nec-
The study conducted by Iftemi et al. (2023) essary modifications.
explored the broader implications and potential
applications of quantum computing within the Methodology
context of cyber-physical systems. The researchers’
investigations provided clarification on the potential This work presents a methodology for integrating
enhancements in capabilities and efficiency of cyber- post-quantum models, facilitated by AI, into quan-
physical systems (CPS) through the utilization of tum cyber-physical systems (QCPS) (Rajawat et al.,
quantum processing. They presented an analysis of 2022).
Vaidyan and Hybrid classical-quantum Effective fault Complexity of hybrid Integration of more
Tyagi, 2022 AI models for fault analysis, potential for models, potential quantum algorithms?
analysis rapid diagnostics scalability issues
Almutairi et al., Quantum dwarf mongoose Enhanced intrusion Possibly high Integration with other
2023 optimization with detection utilizes computational intrusion detection
ensemble deep learning for quantum optimization overhead mechanisms?
intrusion detection
Kobayashi et al., Fully automated data Full automation of Limited to laser Automation in other
2021 acquisition for laser data acquisition, production domain, domains of CPS?
production CPS Potential for higher hardware restrictions?
precision
Zhu et al., 2023 Learning spatial graph Scalability , effective Might require vast Other applications
structure for KPI anomaly anomaly detection for amounts of training of the spatial graph
detection in large-scale KPIs data model?
CPS
Applied Data Science and Smart Systems 493
Identify application areas – Identify the specific sce- It is feasible to create a function F that integrates the
narios in which the integration of AI with post- components (C, A, P) and maps them to an output,
quantum models might contribute significantly which represents the performance or efficiency of the
to quantum computing problem-solving (QCPS). integrated system.
Concentrate your developmental endeavors on
those areas (Iftemi et al., 2023). Potential areas QCPS = F(C, A, P) (1)
of focus include secure communication, decen-
tralized management, and real-time optimiza-
The extraction of sub-functions that represent
tion.
interactions between components can be performed
Select appropriate AI and post-quantum algorithms –
on the function F. The optimization of a CPS’s effi-
It is imperative to exercise careful consideration
ciency (Zhu et al., 2023) can be achieved through the
while selecting AI algorithms in order to ensure
utilization of an AI model, denoted as function f1.
their ability to effectively address the issues in-
f1(A, C) = performance improvement of CPS using AI.
herent in the quantum computing for public
The enhancement of CPS security, denoted as f2,
safety (QCPS) (Vereno et al., 2023) scenario. In
can be further augmented by the utilization of post-
a comparable manner, select post-quantum cryp-
quantum cryptography techniques.
tography algorithms that exhibit both robust se-
f2(P, C) = security enhancement of CPS using post-
curity and sufficient efficiency for their intended
quantum cryptography.
applications.
It is feasible to represent the performance of the
Develop integrated AI-PQ modules – There is a need
integrated system by aggregating the individual
to create and develop modules that integrate AI
components.
and post-quantum cryptography characteristics.
The optimization of these components is nec-
essary to minimize resource consumption and QCPS = F(C, A, P) = g(f1(A, C), f2(P, C)) (2)
provide seamless integration into the existing cy-
ber-physical systems (CPS) (Vaidyan and Tyagi, The function g incorporates considerations of both
2022) network. enhanced efficiency and heightened safety (Li et al.,
Implement AI-PQ modules in QCPS – It is impera- 2018).
tive to ensure compatibility with current hard- Through a comprehensive examination of the
ware and software when integrating the AI-PQ characteristics exhibited by F and its subordinate
modules into the QCPS architecture. Modifica- functions, a deeper understanding can be obtained
tions to elements such as data formats, control regarding the advantages and disadvantages associ-
systems, and communication protocols may po- ated with the utilization of post-quantum models
tentially be needed. facilitated by artificial intelligence in the context of
Evaluate performance and security – This analysis quantum computing for problem-solving. The dif-
aims to evaluate the level of integration and safe- ficulty of integrating artificial intelligence and post-
ty of the AI-PQ modules within the QCPS infra- quantum cryptography into cyber-physical systems
structure. It is imperative to analyze the impact (CPS) can be assessed by examining the complexity of
of a given factor on latency, throughput, and se- the function F (Tangsuknirundorn et al., 2017). The
curity (Almutairi et al., 2023). function F in QCPS is subject to constraints on the
Refine and iterate – Enhance the components of AI- available resources, which are represented by inputs
PQ and their integration into the QCPS based on C, A, and P. The challenges associated with evaluating
the evaluation outcomes. This iterative method and enhancing the performance of function F can be
ensures consistent progress and adherence to seen as an apt analogy for the obstacles faced in the
evolving requirements. processes of verification and validation.
Mathematical models can undergo analysis and
The utilization of AI in conjunction with post- optimization to identify strategies for integrating
quantum cryptography (PQ) within the realm of AI-enabled post-quantum models into quantum com-
cyber-physical systems (CPS) enables the development puting problem solving (QCPS) systems, effectively
of mathematical models that effectively capture the leveraging the former while minimizing the impact of
intricate relationships and interdependencies among the later. Consequently, we are potentially approach-
these components. ing a pivotal moment characterized by a technologi-
Consider a system including of AI (Kobayashi et cal revolution, wherein the development of quantum
al., 2021) models represented as A, a collection of cyber-physical systems that include attributes of secu-
cyber components represented as C, and a set of post- rity, efficiency, and intelligence is underway (Yevseiev
quantum cryptography algorithms represented as P. et al., 2022).
494 Integrating AI-enabled post-quantum models in quantum cyber-physical systems
Table 63.2 Datasets relevant to quantum cyber-physical systems (QCPS) that incorporate AI-enabled post-quantum model
integration
Quantum dataset 1 Data simulating quantum effects Quantum computing Q lab research
in CPS simulation
AI quantum dataset 2 Dataset for AI algorithms on AI quantum integration AI cyber quantum institute
quantum data
PQ protocols 3 Post-quantum cryptographic Post-quantum cryptography PQ crypto foundation
protocol simulations
CPS Real World 4 Real-world CPS data integrated Quantum CPS real-world CPSNet research
with quantum computing application
QCPS test bench 5 Benchmark dataset for QCPS Performance testing QCPS global consortium
systems performance
- Detect noise in quantum system and correct or Resource limitations – The implementation of AI
adjust using AI. and post-quantum cryptography algorithms on
2. Synchronization: quantum computing platforms (QCPS) may en-
- Ensure quantum computations, AI predictions, counter challenges arising from limited process-
and CPS operations are well synchronized. ing resources, memory capacity, and energy lim-
3. Scalability: its, thereby hindering their efficient execution.
- Handle growth in system components, data, Verification and validation – The implementation
and computational requirements. of verification and validation methods for AI-
4. Interoperability: enabled post-quantum models in quantum com-
- Ensure seamless interaction between AI, PQ, puting and post-quantum cryptographic systems
and CPS components. (QCPS) might pose challenges in terms of time
5. PostQuantumCryptoOverhead: consumption and complexity. However, these
- Manage time and resource overhead intro- procedures are crucial for guaranteeing the ac-
duced by PQ encryption/decryption. curacy, security, and reliability of the models
End Module (Khoshnoud et al., 2017).
The provided code presents a theoretical perspec- Notwithstanding these challenges, the integration
tive (Niemann et al., 2021) on the possible interac- of AI-enabled post-quantum models in quantum
tion among artificial intelligence, post-quantum, and computing and physical systems (QCPS) has signifi-
cyber-physical systems inside a quantum environ- cant promise for revolutionizing human interactions
ment. The specific requirements would be contingent and management of the physical environment. This
upon the hardware, software, and domain-specific has the potential to yield innovative advancements
demands (Zajac and Störl, 2022). in secure, intelligent, and interconnected technology
(Figure 63.1).
Opportunities
Opportunities
Enhanced security – The utilization of post-quantum Enhanced security – The use of post-quantum cryp-
cryptography techniques enables the achieve- tographic protocols in quantum CPS can offer
ment of secure long-term storage and transmis- enhanced security against quantum attacks.
sion of private information within a quantum Optimized performance – AI can optimize the perfor-
computing protection system (QCPS), thereby mance of quantum CPS by providing intelligent
mitigating the risks posed by quantum comput- decision-making and predictive maintenance.
ing threats. Resilience and adaptability – AI and PQ integra-
Improved performance – The utilization of AI models tion may lead to systems that can adapt to new
has the potential to enhance performance and threats and continue to operate under adverse
efficiency in quality control and production sys- conditions.
tems (QCPS) by optimizing resource allocation, Innovative applications – This integration could open
control methodologies, and decision-making new avenues for innovative applications in vari-
processes. ous sectors such as healthcare, transportation,
New applications – The integration of AI with post- and smart cities.
quantum cryptography (PQC) has the potential
to enable novel uses of quantum computing and Challenges
post-quantum secure (QCPS) systems. These ap- Complexity of integration – Combining AI, PQ, and
plications include the establishment of secure CPS requires handling complex and possibly
quantum communication networks, the develop- conflicting requirements.
ment of autonomous quantum control systems, Quantum decoherence – The instability of quantum
and the realization of real-time quantum optimi- states can pose challenges in maintaining consis-
zation. tent quantum computation for CPS (Ahmad et
al., 2021).
Challenges
Scalability – Post-quantum cryptographic methods
Integration complexity – The integration of AI and may introduce significant overhead, which can
post-quantum cryptography (PQC) into cur- be a challenge for scalable quantum CPS.
rent cyber-physical systems (CPS) infrastructures AI Interpretability – AI decision-making processes
might pose challenges due to factors such as need to be transparent, especially in critical cy-
compatibility, resource constraints, and the im- ber-physical systems where errors can have se-
perative for real-time performance. vere consequences.
496 Integrating AI-enabled post-quantum models in quantum cyber-physical systems
Figure 63.1 Integrating AI-enabled post-quantum models in quantum cyber-physical systems opportunities and chal-
lenges
AI model complexity Number of layers, neurons, etc., in 10 layers, 1000 Affects computation time
the AI model neurons
Quantum bits (Qubits) Number of qubits in the quantum 50 qubits Defines quantum capacity
system
PQ algorithm Post-quantum algorithm used NTRU, Kyber, etc. Affects security & performance
CPS network size Number of devices/nodes in the 100 nodes Affects network scalability
CPS network
CPS update frequency How often the CPS updates its Every 10 ms Affects system responsiveness
state/data
Noise level Level of noise in the quantum 0.01% Impacts quantum reliability
system
AI training data size Amount of data used for training 10 GB Affects AI accuracy
the AI model
PQ key size Size of the cryptographic keys used 2048 bits Balances security & speed
Quantum gate depth Depth of quantum circuits 500 gates Affects quantum computation
(number of gates in sequence)
AI inference speed Time taken for the AI model to 50 ms per input Affects real-time decision
process input and produce output
AI model complexity 15 layers, 1500 neurons Slight increase in accuracy but Complexity trade-off to be
higher computational cost considered
Quantum bits (Qubits) 60 qubits Enhanced quantum processing Error correction techniques
capability but more noise needed
PQ algorithm Kyber Secure communication but moderate Suitable for medium-
computational overhead security tasks
CPS network size 150 nodes Increased network delay, but better Scalability concerns arise
distributed processing
CPS update frequency Every 5 ms More real-time updates, but higher Need efficient data
bandwidth consumption transmission
Noise level 0.02% Slight degradation in quantum Requires better noise
computations isolation
AI training data size 12 GB Improved model accuracy by 2% Diminishing returns beyond
10 GB
PQ key size 3072 bits Enhanced security but longer key Key size to be chosen based
generation time on needs
Quantum gate depth 600 gates Extended computational possibilities Deep circuits need error
but more errors mitigation
AI inference speed 40ms per input Faster real-time decision-making Optimal for time-sensitive
tasks
In the absence of empirical simulation outcomes, previously inconceivable. Additionally, the inclusion
this discussion will outline the potential transforma- of AI components further enhances CPS by providing
tion of the previously mentioned “Simulation param- intelligent analysis, adaptability, and decision-making
eter table” into a “Results analysis” table, which capabilities. The integration of various systems such
pertains to the integration of AI-enabled post-quan- as healthcare, transportation, and energy infrastruc-
tum models into quantum cyber-physical systems ture could potentially yield a more responsive, secure,
(CPS). and efficient outcome.
Table 63.4 shows the simulated impact of altering However, these advancements are not devoid of
specific settings from their default values. The retrieval challenges. Comprehensive research is essential in
of actual values from the simulation is necessary, and order to ascertain the most effective methods for
any comments, observations, and impacts should be ensuring dependable and secure interactions among
based on empirical data and analysis conducted in a AI, post-quantum (PQ) systems, and quantum com-
real-world context. ponents. The concerns encompass quantum noise,
potential security vulnerabilities in artificial intelli-
Conclusion gence, and the nascent state of post-quantum cryp-
tography methodologies. Ultimately, the successful
The integration of quantum cyber-physical systems incorporation of AI-enabled post-quantum models
(CPS) with AI-enabled post-quantum (QCPS) models into quantum cyber-physical systems (CPS) holds
represents a significant and transformative conver- great promise for the future, offering a multitude of
gence of advanced technologies. We are currently at potential opportunities. However, achieving this goal
the threshold of a forthcoming era in the design and will necessitate thorough investigation, robust design
operation of cyber-physical systems (CPS). This age methodologies, and collaborative efforts across vari-
entails the integration of AI, which possesses the abil- ous disciplines. The road is in its early stages, but it
ity to make predictions, with the strong cryptographic holds the potential to catalyze a transformative shift
capabilities offered by post-quantum mechanisms, as in the realm of cyber-physical systems.
well as the immense processing power provided by
quantum systems.
The quantum aspect of cyber-physical systems
References
(CPS) offers a multitude of opportunities, enabling Tosh, D., Galindo, O., Kreinovich, V., and Kosheleva, O.
enhanced performance and functionalities that were (2020). Towards security of cyber-physical systems us-
498 Integrating AI-enabled post-quantum models in quantum cyber-physical systems
ing quantum computing algorithms. 2020 IEEE 15th Tangsuknirundorn, P., Sooraksa, P., and Sooraksa, P.
Int. Conf. Sys. Sys. Engg. (SoSE), 313–320. (2017). Design of a cyber-physical demonstration us-
Zhang, Xin, Zhao Yang Dong, Zeya Wang, Chixin Xiao, ing STEAM: Superconducting chaotic robots. 2017
and Fengji Luo. (2015). Quantum cryptography based 21st Int. Comp. Sci. Engg. Conf. (ICSEC), 1–5.
cyber-physical security technology for smart grids. Yevseiev, S., Milevskyi, S., Bortnik, L., Alexey, V., Bonda-
51–6, DOI: 10.1049/ic.2015.0263. renko, K., and Pohasii, S. (2022). Socio-cyber-physical
Rajawat, A. S., Goyal, S. B., Bedi, P., Constantin, N. B., systems security concept. 2022 Int. Cong. Hum.-
Raboaca, M. S., and Verma, C. (2022). Cyber-physical Comp. Interac. Optim. Rob. Appl. (HORA), 1–8.
system for industrial automation using quantum deep Mekala, M. S., Srivastava, G., Gandomi, A. H., Park, J.
learning. 2022 11th Int. Conf. Sys. Model. Adv. Res. H., and Jung, H.-Y. (2023). A quantum-inspired sen-
Trends (SMART), 897–903. sor consolidation measurement approach for cyber-
Iftemi, A., Cernian, A., and Moisescu, M. A. (2023). Quan- physical systems. IEEE Trans. Netw. Sci. Engg., 1–14.
tum computing applications and impact for cyber doi:10.1109/tnse.2023.3301402.
physical systems. 2023 24th Int. Conf. Con. Sys. Lu, K.-D. and Wu, Z.-H. (2022). Genetic algorithm-based
Comp. Sci. (CSCS), 377–382. cumulative sum method for jamming attack detec-
Vereno, D., Khodaei, A., Neureiter, C., and Lehnhoff, S. tion of cyber-physical power systems. IEEE Trans. In-
(2023). Exploiting quantum power flow in smart grid strum. Meas., 71, 1–10.
co-simulation. 2023 11th Workshop Model. Simul. Singh, J., Singh, S., Singh, S., and Singh, H. (2019). Evaluat-
Cyber-Phy. Ener. Sys. (MSCPES), 1–6. ing the performance of map matching algorithms for
Vaidyan, V. M. and Tyagi, A. (2022). Hybrid classical-quan- navigation systems: an empirical study. Spat. Inform.
tum artificial intelligence models for electromagnetic Res., 27, 63–74.
control system processor fault analysis. 2022 IEEE Asif, R. and Buchanan, W. J. (2017). Seamless crypto-
IAS Glob. Conf. Emerg. Technol. (GlobConET), 798– graphic key generation via off-the-shelf telecommu-
803. nication components for end-to-end data encryption.
Almutairi, Laila, Ravuri Daniel, Shaik Khasimbee, E. Lax- 2017 IEEE Int. Conf. Internet of Things (iThings)
mi Lydia, Srijana Acharya, and Hyunil Kim. (2023). IEEE Green Comput. Comm. (GreenCom) IEEE Cy-
Quantum Dwarf Mongoose Optimization with En- ber Phy. Soc. Comput. (CPSCom) IEEE Smart Data
semble Deep Learning Based Intrusion Detection in (SmartData), 910–916.
Cyber-Physical Systems. IEEE Access, 11, 66828– Niemann, P., Mueller, L., and Drechsler, R. (2021). Combin-
66837. ing SWAPs and remote CNOT gates for quantum cir-
Kobayashi, Y., Takahashi, T., Nakazato, T., Sakurai, H., cuit transformation. 2021 24th Euromicro Conf. Dig.
Tamaru, H., Ishikawa, K. L., Sakaue, K., and Tani, Sys. Des. (DSD), 495–501.
S. (2021). Fully automated data acquisition for laser Zajac, M. and Störl, U. (2022). Towards quantum-based
production cyber-physical system. IEEE J. Sel. Top. search for industrial data-driven services. 2022 IEEE
Quan. Elec., 27(6), 1–8. Int. Conf. Quan. Softw. (QSW), 38–40.
Zhu, Haiqi, Seungmin Rho, Shaohui Liu, and Feng Ji- Khoshnoud, F., de Silva, C. W., and Esat, I. I. (2017). Quan-
ang. (2023). Learning Spatial Graph Structure for tum entanglement of autonomous vehicles for cyber-
Multivariate KPI Anomaly Detection in Large-scale physical security. 2017 IEEE Int. Conf. Sys. Man Cy-
Cyber-Physical Systems. IEEE Transactions on In- bernet. (SMC), 2655–2660.
strumentation and Measurement, 72, DOI: 10.1109/ Ahmad, S. F., Ferjani, M. Y., and Kasliwal, K. (2021). En-
TIM.2023.3284920. hancing security in the industrial IoT sector using
Li, S., Ni, Q., Sun, Y., Min, G., and Al-Rubaye, S. (2018). quantum computing. 2021 28th IEEE Int. Conf. Elec.
Energy-efficient resource allocation for industrial cy- Cir. Sys. (ICECS), 1–5.
ber-physical IoT systems in 5G era. IEEE Trans. In-
dus. Inform., 14(6), 2618–2628.
64 Adaptive resource allocation and optimization in cloud
environments: Leveraging machine learning for efficient
computing
Anand Singh Rajawat1, S. B. Goyal2,a, Manoj Kumar3 and Varun Malik4
1
School of Computer Science & Engineering, Sandip University, Nashik, Maharashtra, India
2
City University, Petaling Jaya, 46100, Malaysia
3
University of Wollongong, Dubai, UAE
4
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
In contemporary cloud computing environments, the efficient allocation and utilization of resources are vital to ensure
prompt performance and maximize the utilization of the existing infrastructure. The proliferation of cloud platforms has led
to the emergence of considerable challenges related to load balancing and efficient task scheduling, since a rising number of
applications and services rely on these platforms. This article introduces an innovative approach to tackle these challenges
through the utilization of machine learning (ML) techniques. In this study, we propose a comprehensive framework that
effectively allocates resources in real-time systems by adapting to their evolving demands. This framework achieves its objec-
tives by integrating algorithms for load balancing and scheduling. Machine learning models, which have been trained using
previous data on workloads and system performance, can be utilized to forecast upcoming load surges and identify potential
bottlenecks. Subsequently, the computer system proactively modifies the allocation of resources and the arrangement of tasks
to preemptively address future challenges. Upon comparison with conventional approaches, the initial findings indicate sig-
nificant advantages in terms of system performance, decreased latency, and improved resource utilization. Furthermore, the
framework’s flexible architecture ensures the capacity to scale and adapt, rendering it well-suited for deployment in dynamic
environments such as cloud-based systems that undergo frequent modifications. This study showcases the transformative
potential of ML in redefining resource allocation and task scheduling inside cloud computing ecosystems.
Keywords: Cloud computing, adaptive resource allocation, load balancing algorithms, scheduling algorithms, machine learn-
ing optimization, efficient computing
a
drsbgoyal@[Link]
500 Adaptive resource allocation and optimization in cloud environments
efficiency. In contrast to the previously employed enhance the energy efficiency of data transmission
static models, the current dynamic and flexible archi- operations, which is a critical necessity in the age of
tecture holds the potential for enhanced processing the Internet of Things (IoT) and ubiquitous wireless
throughput, reduced latency, and increased overall communications.
system performance. Kumari and Saxena (2021) proposed the imple-
This study investigates the possible impact of mentation of an “Advanced fusion ACO approach”
machine learning on cloud-based resource manage- as a means to enhance the optimization of memory in
ment and optimization. Our objective is to provide cloud computing. The methodology utilized in their
insights into potential avenues for enhancing the study involves the implementation of the ant colony
efficiency and adaptability of cloud computing. This optimization (ACO) algorithm, which draws inspira-
will be achieved by a comprehensive examination of tion from the behavior of ants. This approach offers
load balancing and scheduling approaches that are a practical solution for mitigating memory inefficien-
augmented using machine learning techniques. Our cies inside cloud systems.
research contributions are as follows: The present study explores the utilization of
dynamic binary translation (DBT) cache inside cloud
• This study introduces a ML-based framework for computing environments. In order to enhance system
efficient resource allocation and task scheduling performance in cloud environments, a group of aca-
in cloud computing environments. demics devised a specialized optimization technique
• The proposed framework predicts load surges for the store and retrieval operations of the DBT
and optimizes resource allocation, outperform- cache (Yi, 2020).
ing traditional approaches in system performance A comprehensive tabular representation of the ref-
and latency reduction. erenced scholarly papers, encompassing sections on
• By integrating ML algorithms, the framework citation, methods, benefits, drawbacks, and future
dynamically adapts to evolving demands, ensur- directions for study (Table 64.1).
ing scalability and adaptability in cloud-based Table 64.1 provides a concise overview of the
systems. various methodologies, highlighting their respec-
tive merits, drawbacks, and potential areas for
The paper organization – The related work, the pro- future investigation (Chaitra et al., 2020). The ben-
posed methodology, results analysis and finally con- efits, downsides, and research gaps are synthesized
clusion and future work. in accordance with the referenced literature; a more
comprehensive examination of each study may be
Related work required to extract subtle nuances.
instances of prevalent scheduling algorithms (Gao et predict the resource requirements of applications and
al., 2020; Singh et al., 2020): workloads in the future. The utilization of this data
First come first served (FCFS) – The tasks are exe- can enhance the efficacy of algorithms employed in
cuted sequentially according to the order in which load balancing and scheduling by facilitating more
they were received by the algorithm. precise determinations pertaining to resource alloca-
Shortest job first (SJF) – The algorithm exhibits a tion (Bilgaiyan et al., 2014).
preference for the task with the shortest duration. Optimize resource allocation decisions – The opti-
Priority scheduling – This algorithm prioritizes mization of resource allocation can be enhanced by
activities based on characteristics such as urgency the utilization of machine learning models, which
and relevancy, executing them in the order of their include both projected resource demands and desired
priority. levels of service quality. Machine learning models can
be effectively utilized in several scenarios, such as
determining the most efficient method for distributing
Algorithm: Adaptive Resource Allocation and traffic over multiple servers or identifying the ideal
Optimization using ML number of servers to allocate for a certain application
Inputs: (Kumar et al., 2017).
- List of tasks: tasks[]
- List of available cloud resources: resources[]
- ML model for load balancing: ML_LoadBalancer Detailed steps for the proposed machine learning
- ML model for task scheduling: ML_Scheduler model
Output: User requests – User requests are funneled in through
- Efficient allocation and execution of tasks this entry point in the cloud environment when they
Procedure: are processed. These could range from easy activities
1. INITIALIZE empty list allocatedTasks[] and like retrieving data to more difficult ones like doing
scheduledTasks[] sophisticated computations.
2. FOR each task in tasks[]: Load balancer – The load balancer disperses
2.1 Predict optimal resource using incoming requests among several cloud servers so as
ML_LoadBalancer to maximize throughput, reduce response time, opti-
resource = ML_LoadBalancer.predict(task) mize resource consumption, and prevent overloading
2.2 IF resource is available: of any one resource in particular.
2.2.1 ALLOCATE task to resource Machine learning model – The machine learning
2.2.2 ADD task to allocatedTasks[] model performs an analysis of historical data to deter-
2.3 ELSE: mine workload patterns, resource utilization, and the
2.3.1 QUEUE task for later allocation effectiveness of scheduling decisions in the past in
3. WHILE allocatedTasks[] is not empty: order to forecast optimal allocation techniques.
3.1 FOR each task in allocatedTasks[]:
3.1.1 Predict optimal execution time using
ML_Scheduler
executionTime = ML_Scheduler.predict(task)
3.1.2 SCHEDULE task based on predicted
executionTime
3.1.3 MOVE task from allocatedTasks[] to
scheduledTasks[]
4. EXECUTE all tasks in scheduledTasks[]
5. RETURN “All tasks executed successfully”
End Algorithm
Dynamic resource allocation – Algorithms that optimization through machine learning presents a
dynamically allocate resources make adjustments to promising and innovative avenue.
those resources in real time, based on forecasts and This study proposes an adaptive resource allocation
the present status of the system, in order to provide and optimization formula for cloud environments,
load balancing across the cloud infrastructure. specifically focusing on load balancing methods and
Task scheduling – The sequence in which activities scheduling algorithms. The formula incorporates
are carried out can be determined by scheduling algo- machine learning techniques to enhance the efficiency
rithms, which take into account dependencies, prior- and effectiveness of resource allocation and optimiza-
ity, and the expected execution timeframes provided tion in cloud settings.
by the ML model (Peng et al., 2016).
Optimized resource allocation – When allocating R(t) = f(W(t), Q(t), M(t))
resources, an optimum strategy is used, taking into
account both the current state of the system and the where:
predictions generated by the machine learning model. R(t) is the resource allocation at time t
This guarantees that jobs are assigned to the resources W(t) is the workload at time t
that are the most appropriate for them. Q(t) is the quality of service requirements at time t
Execution on cloud resources – The tasks are car- M(t) is the machine learning model at time t
ried out on the resources provided by the cloud in Through the analysis of historical data, machine
accordance with the optimum allocation and time- learning models can acquire the ability to compre-
table in an effort to achieve high levels of efficiency hend the intricate relationship between workload,
while keeping operational costs to a minimum. service quality requirements, and resource alloca-
Monitoring – The execution of the tasks and the tion. After the completion of the training process,
use of the resources are continuously monitored by the model can be employed to ascertain the optimal
the system in order to give real-time data for the allocation of resources in order to fulfill service level
machine learning model. agreements (SLAs) for various workloads and levels
of service quality.
Feedback data – The ML model receives feedback
The machine learning model can ensure its align-
in the form of performance data and the outcomes
ment with fluctuations in workload and adherence to
of the resource allocation and task executions. This
service quality standards. The outcome is a heightened
enables the model to learn and adapt over time, which
level of adaptability and efficiency in the allocation of
improves its ability to make predictions and judg-
existing resources (Nuradis and Lemma, 2019).
ments in the future.
The utilization of machine learning models can
enhance both load balancing and scheduling deci-
Benefits of adaptive resource allocation and optimiza-
sions. The model can be utilized, for example, to
tion
determine the most efficient method of distributing
There exist multiple favorable consequences that can
traffic among multiple servers or the optimal alloca-
arise from the adaptive allocation and optimization
tion of servers for a particular application. The model
of resources.
has the ability to be utilized for the purpose of priori-
Improved performance – The utilization of adap-
tizing server operations and determining the optimal
tive resource allocation and optimization techniques
order in which jobs should be executed.
can improve the performance of cloud applications The enhancement of performance, cost-effective-
and workloads. ness, dependability, and scalability in cloud com-
Reduced costs – The avoidance of overprovision- puting environments can be achieved through the
ing of resources is achieved by the implementation of utilization of adaptive resource allocation and opti-
adaptive resource allocation and optimization tech- mization techniques leveraging machine learning.
niques, which effectively contribute to the mainte- The subsequent representation presents a potential
nance of affordable cloud costs. instantiation of the formula for distributing resources
Increased reliability – The enhancement of the to a cloud-based web application.
dependability of cloud applications and workloads The machine learning model can be trained using
can be achieved through the implementation of adap- historical data in order to uncover the relationship
tive resource allocation and optimization techniques, between the number of users of a web application,
which aim to eliminate any limitations on available the response time required, and the optimal number
resources (Peng et al., 2016). of web servers. Once the model has undergone train-
In order to enhance efficiency, economy, reliability, ing, it can be utilized to ascertain the number of web
and scalability within the realm of cloud computing, servers necessary to meet the response time demands
the utilization of adaptive resource allocation and of a specific user demographic.
504 Adaptive resource allocation and optimization in cloud environments
The machine learning model has the potential the quality of service, we utilize the quality-of-service
to be regularly updated in order to accurately cap- violation as a means of verification.
ture the latest requirements pertaining to through- To ensure accurate resource allocation of the
put and reaction time. The allocation of resources model, the inclusion of the following restrictions is
would exhibit more flexibility and efficacy as a recommended:
consequence. R ij ∈{0,1}, where R ij is 1 if the $i$th IoT device
The enhancement of performance (Shin, 2014), is deployed on the $j$th edge server and 0 otherwise
cost-effectiveness, dependability, and scalability in Σ jM R ij =1 for all i
cloud computing environments can be achieved Σ iN R ij ≤Cmax for all j
through the utilization of adaptive resource allocation where C max is the maximum capacity of an edge
and optimization techniques that leverage machine server.
learning. The problem at hand can be addressed by the utili-
The subsequent illustration presents a mathemati- zation of several machine learning techniques, includ-
cal model that could be employed in tandem with ing linear programming, integer programming, and
machine learning techniques to facilitate adaptive reinforcement learning.
resource allocation and optimization inside cloud Leveraging machine learning for efficient
environments. computing.
minimize: There exist multiple methodologies via which
machine learning can be employed to enhance the
efficiency of resource allocation and optimization.
J(R) = Σ_i^N w_i * d_i(R) + Σ_j^M c_j * C_j(R)
One notable application of machine learning are as
+ Σ_k^K v_k * V_k(R) follows:
Predict future resource needs – The ability to fore-
where: cast the future resource demands of IoT applications
R is the resource allocation can be achieved through the utilization of machine
N is the number of IoT devices learning models that have been trained on historical
M is the number of IoT applications data. Based on the available facts, we can make more
K is the number of quality of service requirements informed decisions on the allocation of our finite
w i is the weight of the $i$th IoT device financial resources.
d i (R) is the distance between the $i$th IoT device Optimize resource allocation decisions – The opti-
and the edge server mization of resource allocation can be enhanced by
c j is the cost of deploying the $j$th IoT application the utilization of machine learning models, which
on an edge server include both projected resource demands and
C j(R) is the number of edge servers required to desired levels of service quality. In order to enhance
deploy the $j$th IoT application on the resource allo- the equitable allocation of traffic among accessible
cation R edge servers, or to ascertain the optimal alloca-
v k is the weight of the $k$th quality of service tion of edge servers for a certain IoT application,
requirement the utilization of a machine learning model may be
V k(R) is the violation of the $k$th quality of ser- employed.
vice requirement on the resource allocation R. Adapt to changes in the environment – To address
The fundamental purpose of the objective func- the integration of emerging IoT devices and the intro-
tion is to minimize a weighted sum of travel times, duction of novel IoT applications, it is possible to peri-
expenditures, and quality of service breaches. The odically update machine learning models within the
prioritization of various IoT devices, applications, cloud environment. This ensures that our resources
and quality of service requirements can be achieved are being utilized efficiently at all times.
through the utilization of weights. The significance Table 64.2 represents the simulation parameters
of service standards may be amplified if they have utilized in a cloud system that employs machine learn-
greater importance to consumers or if they are asso- ing techniques to achieve optimal resource allocation
ciated with critical or possibly more profitable IoT and optimization.
applications. Table 64.3 provides can be modified according to
Increasing the physical separation between the the specified needs, which may entail making adjust-
edge server and the IoT device might lead to a reduc- ments to the default values or including/excluding col-
tion in latency for the IoT application. One potential umns as necessary. However, depending on the nature
strategy for reducing cloud expenses is by deploying of the cloud environment and the desired outcomes
IoT software on an edge server. In order to ensure of the simulation, certain qualities listed in the table
the fulfillment of consumers’ expectations regarding may hold greater significance than others in practical
Applied Data Science and Smart Systems 505
Table 64.2 simulation parameters utilized in a cloud system that employs machine learning techniques to achieve optimal
resource allocation and optimization.
Environment parameters
Total resources Number of virtual machines or 50–500 100
physical servers
Resource capacity CPU, RAM, storage capacity of Varies (e.g., 2–16 CPUs, 4 CPUs, 8 GB RAM, 200
each resource 4–64 GB RAM, 100–500 GB storage
GB storage)
Workload type Nature of incoming tasks (CPU- CPU-intensive, memory- CPU-intensive
intensive, I/O-intensive, etc.) intensive, balanced, etc.
Task arrival rate Average number of tasks arriving e.g., 10–100 tasks/min 50 tasks/min
per time unit
Task length Duration to complete a task Varies based on workload 5 minutes
type
Load balancing algorithms
Algorithm type Type of load balancing algorithm Round Robin, least Round Robin
used connection, weighted
distribution, etc.
Prediction window In case of predictive algorithms, 5–15 minutes 10 minutes
the time window for prediction
Scheduling algorithms
Algorithm type Type of scheduling algorithm used First come first serve FCFS
(FCFS), shortest job first
(SJF), priority-based, etc.
Preemption Ability to interrupt a currently Enabled/disabled Disabled
running task for a higher priority
one
Machine learning parameters
ML algorithm Machine learning algorithm used Decision trees, neural Neural networks
for prediction/optimization networks, SVM, etc.
Training data size Number of past data points used 10,000–100,000 data points 50,000 data points
for training the model
Features Features/parameters considered by Resource utilization, task Resource utilization, task
the ML model length, arrival rate, etc. length
Model update Frequency at which the ML model After every 1000 tasks, 24 After every 1000 tasks
frequency is updated/retrained hours, etc.
Performance metrics
Throughput Number of tasks completed per Tasks per minute/hour -
time unit
Resource utilization Percentage of resources being used 0–100% -
Waiting time Time a task waits before it starts Time units (e.g., seconds, -
executing minutes)
Makespan Total time taken to complete all Time units (e.g., minutes, -
tasks hours)
application. The analysis of results table can be uti- Achieving consistent outcomes can be facilitated
lized to present the final outcomes of the simulation through the integration of Round Robin scheduling,
across different configurations. As the availability of first-come-first-serve (FCFS) scheduling, and neural
actual numerical data is lacking, it is important to networks. The least connection algorithm exhibited
note that the table provided serves solely as an illus- reduced waiting times compared to the baseline, while
trative sample. Once the simulations have been con- achieving equivalent throughput. The implementation
cluded, the actual findings can be inputted. of shortest job first (SJF) scheduling has resulted in a
506 Adaptive resource allocation and optimization in cloud environments
Table 64.3 Results analysis
decrease in wait times while maintaining a high level integration of weighted distribution, a priority-based
of throughput, hence enhancing overall efficiency. The scheduler, and support vector machines (SVM). There
attainment of optimal performance is realized by the is a lack of significant advancement. In comparison
Applied Data Science and Smart Systems 507
to the control group, the performance of SVM with management systems capable of acquiring knowledge
Round Robin and FCFS did not exhibit statistically and adjusting to novel workloads without the need
significant improvement. Several key observations for human intervention. Utilizing machine learning
can be derived from the presented tabular data: techniques for the purpose of dynamic resource man-
The achievement of optimal overall performance is agement and optimization in cloud-based environ-
facilitated by the utilization of a weighted distribution ments is not only a prominent advancement in the
strategy in combination with a priority-based sched- realm of efficient computing, but also an impera-
uler and support vector machines (SVM). Machine tive necessity. The integration of machine learning is
learning techniques, such as support vector machines expected to play a crucial role in ensuring the effi-
(SVMs), demonstrate exceptional performance in cacy, efficiency, and reliability of cloud operations as
certain scenarios, while encountering challenges in the cloud ecosystem progresses and reaches a more
alternative contexts. The duration of waiting periods advanced stage.
for tasks is significantly impacted by the scheduling
process. Drawing more accurate conclusions can be References
achieved by analyzing the real data presented in the
results analysis table subsequent to conducting the Luo, G., Zhang, Z., Wu, D., Li, X., and Liu, G. (2017).
Research on optimization allocation of manufactur-
simulation using the actual settings and noting the
ing resource services in the cloud environment. 2017
observed outcomes. IEEE SmartWorld Ubiquit. Intel. Comput. Adv. Trust.
Comput. Scal. Comput. Comm. Cloud Big Data Com-
Conclusion put. Inter. People Smart City Innov. (SmartWorld/
SCALCOM/UIC/ATC/CBDCom/IOP/SCI), 1–6.
To efficiently handle substantial volumes of data and Raj, H., Ojha, S. K., and Nazarov, A. (2020). A hybrid ap-
accommodate dynamic user requirements, the field proach for process scheduling in cloud environment
of cloud computing has seen advancements necessi- using particle swarm optimization technique. 2020
tating the implementation of increasingly advanced Int. Conf. Engg. Telecomm. (En&T), 1–5.
approaches for resource allocation and optimiza- Hengbo, X. and Yu, C. (2022). Energy consumption optimi-
tion. Although classic load balancing and scheduling zation method for wireless communication data trans-
mission in cloud environment. 2022 IEEE Int. Conf.
approaches have been widely used, they can prove
Artif. Intel. Comp. Appl. (ICAICA), 228–231.
inadequate in the dynamic and extensive cloud sys- Kumari, P. and Saxena, A. S. (2021). Advanced fusion ACO
tems prevalent in contemporary times. The integra- approach for memory optimization in cloud comput-
tion of these methodologies with machine learning ing environment. 2021 Third Int. Conf. Intel. Comm.
has the potential to serve as an effective technique for Technol. Virt. Mob. Netw. (ICICV), 168–172.
surmounting these challenges. The optimization of Yi, D. (2021). Dynamic binary translation cache optimiza-
cloud resource utilization mostly relies on adaptive tion algorithm in cloud computing environment. 2021
resource allocation techniques, encompassing load Glob. Reliab. Progn. Health Manag. (PHM-Nanjing),
balancing and scheduling algorithms. The primary 1–5.
goal is to achieve a balanced equilibrium in resource Chaitra, T., Agrawal, S., Jijo, J., and Arya, A. (2020). Multi-
allocation, ensuring that resources are neither unde- objective optimization for dynamic resource pro-
visioning in a multi-cloud environment using lion
rutilized nor excessively strained, while simultane-
optimization algorithm. 2020 IEEE 20th Int. Symp.
ously optimizing throughput and minimizing latency. Comput. Intel. Informat. (CINTI), 000083–000090.
The utilization of machine learning techniques, with Kumar, N. (2023). Spider monkey optimization based re-
their predictive capabilities and reliance on data- source provisioning in cloud computing environment.
driven methodologies, has played a pivotal role in 2023 10th Int. Conf. Sig. Proc. Integr. Netw. (SPIN),
enhancing the adaptability of resource allocation 121–125.
methodologies. Machine learning algorithms have Valarmathi, R. and Sheela, T. (2017). A comprehensive sur-
the capability to enhance the distribution of work- vey on task scheduling for parallel workloads based
loads and resource management through the use of on particle swarm optimization under cloud environ-
historical consumption patterns and real-time data ment. 2017 2nd Int. Conf. Comput. Comm. Technol.
analysis. By implementing measures to ensure timely (ICCCT), 81–86.
Yang, Z., Liu, M., Xiu, J., and Liu, C. (202). Study on cloud
completion of tasks, the efficiency of the cloud sys-
resource allocation strategy based on particle swarm
tem can be enhanced, hence positively impacting the ant colony optimization algorithm. 2012 IEEE 2nd
overall user experience. In addition, the implementa- Int. Conf. Cloud Comput. Intel. Sys., 1, 488–491.
tion of a self-adaptive system powered by machine Huang, Z. (2021). Application of artificial intelligence sys-
learning will be of utmost importance as cloud envi- tem in smart education in cloud environment with
ronments become increasingly complex. This facili- optimization models. 2021 5th Int. Conf. Comput.
tates the development of fully autonomous cloud Methodol. Comm. (ICCMC), 313–316.
508 Adaptive resource allocation and optimization in cloud environments
Ga˛sior, J. and Seredyński, F. (2021). An automata-based plications in cloud systems. 2017 3rd Int. Conf. Com-
profit optimization of cloud brokers in IaaS environ- put. Intel. Comm. Technol. (CICT), 1–6.
ment. 2021 IEEE 14th Int. Conf. Cloud Comput. Peng, J., Chen, J., Kong, S., Liu, D., and Qiu, M. Resource
(CLOUD), 723–725. optimization strategy for CPU intensive applications
Mulge, Md Y. and Venkatesh Sharma, K. (2018). Orthogo- in cloud computing environment. 2016 IEEE 3rd
nal Taguchi-based grey wolf optimization algorithm Int. Conf. Cyber Sec. Cloud Comput. (CSCloud),
for task scheduling in cloud environment. 2018 Int. 124–128.
Conf. Elec. Electron. Comm. Comp. Optim. Techniq. Nuradis, J. and Lemma, F. (2019). Hybrid bat and genetic
(ICEECCOT), 1749–1753. algorthim approach for cost effective SaaS placement
Wu, D. (2018). Cloud computing task scheduling policy in cloud environment. 2019 Third Int. Conf. I-SMAC,
based on improved particle swarm optimization. 2018 1–6.
Int. Conf. Virt. Real. Intel. Sys. (ICVRIS), 99–101. Shin, Y.-R. (2014). Optimization for reasonable service price
Pan, K. and Chen, J. (2015). Load balancing in cloud com- in broker based cloud service environment. Fourth Ed.
puting environment based on an improved particle Int. Conf. Innov. Comput. Technol. (INTECH 2014),
swarm optimization. 2015 6th IEEE Int. Conf. Softw. 115–119.
Engg. Ser. Sci. (ICSESS), 595–598. Pattanaik, P. A., Roy, S., and Pattnaik, P. K. (2015). Per-
Gao, M., Zhu, Y., and Sun, J. (2020). The multi-objective formance study of some dynamic load balancing al-
cloud tasks scheduling based on hybrid particle swarm gorithms in cloud computing environment. 2015 2nd
optimization. 2020 Eighth Int. Conf. Adv. Cloud Big Int. Conf. Sig. Proc. Integr. Netw. (SPIN), 619–624.
Data (CBD), 1–5. Nandina, V., Luna, J. M., Lamb, C. C., Heileman, G. L.,
Singh, S., Singh, J., and Sehra, S. S. (2020). Genetic-inspired and Abdallah, C. T. (2014). Provisioning security and
map matching algorithm for real-time GPS trajecto- performance optimization for dynamic cloud envi-
ries. Arab. J. Sci. Engg., 45(4), 2587–2603. ronments. 2014 IEEE 7th Int. Conf. Cloud Comput.,
Yahyaoui, H. and Moalla, S. (2016). CloudFC: Files clus- 979–981.
tering for storage space optimization in clouds. 2016 Sharma, R. and Bharti, M. (2014). Mapping of tasks to re-
IEEE Int. Conf. Cloud Comput. Technol. Sci. (Cloud- sources maintaining fairness using swarm optimiza-
Com), 193–197. tion in cloud environment. Proc. 3rd Int. Conf. Reliab.
Bilgaiyan, S., Sagnika, S., and Das, M. (2014). Workflow Infocom Technol. Optim., 1–6.
scheduling in cloud computing environment using cat Muneotmo, M. and Abe, T. (2015). Designing a distributed
swarm optimization. 2014 IEEE Int. Adv. Comput. design exploration framework in the inter-cloud envi-
Conf. (IACC), 680–685. ronment. 2015 IEEE 8th Int. Conf. Cloud Comput.,
Kumar, B., Kalra, M., and Singh, P. (2017). Discrete binary 1073–1076.
cat swarm optimization for scheduling workflow ap-
65 Quantum deep learning on driven trust-based routing
framework for IoT in the metaverse context
S. B. Goyal1,a, Anand Singh Rajawat2, Jaiteg Singh3 and Chawki Djeddi4
1
City University, Petaling Jaya, 46100, Malaysia
2
School of Computer Science & Engineering, Sandip University, Nashik, Maharashtra, India
3
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
4
Department of Mathematics and Computer Science, Larbi Tebessi University, Tebessa, Algeria
4
LITIS Lab, Rouen University, Rouen, France
Abstract
As the Internet of Things (IoT) progresses towards the concept of the metaverse, it becomes evident that a complex network
of interconnected devices and services emerge, hence demanding innovative strategies for routing and ensuring security. This
study presents a novel architecture that utilizes quantum deep learning techniques to construct a trust-based routing mecha-
nism for IoT landscape within the metaverse. The architecture combines the variational quantum eigensolver (VQE) and
quantum annealing (QA) to achieve this objective. The VQE algorithm, commonly utilized for the purpose of determining
the lowest energy state of quantum systems, is being applied in this study to model trust levels. These trust levels are based
on the historical and real-time interactions of devices. Simultaneously, the proficient professionals at quality assurance (QA)
employ these confidence levels to dynamically construct ideal routing paths. The integration of quantum algorithms and
conventional deep learning methods has dual benefits of safeguarding data privacy and enhancing routing efficiency. This
integration enables the system to efficiently tackle the unique issues presented by the expansive digital environment known
as the metaverse. Based on the first data, it is anticipated that there will be a significant decrease in malicious routing at-
tempts, enhanced throughput, and increased network robustness. The findings presented in this study illustrate the potential
of quantum deep learning to significantly transform future practices in metaverse IoT routing.
Keywords: Quantum machine learning, variational quantum eigensolver (VQE), quantum annealing (QA), trust-based rout-
ing, Internet of Things (IoT), metaverse infrastructure
a
drsbgoyal@[Link]
510 Quantum deep learning on driven trust-based routing framework for IoT in the metaverse context
it is important to collect pertinent data regarding optimization algorithm on IoT devices. The algo-
the decision-making process employed by devices in rithms have the potential to be deployed as a cloud-
selecting data routes, and subsequently analyze the hosted web service. An alternate approach involves
implications of such routing decisions on the overall executing the algorithms on local devices situated at
functionality and performance of IoT network. the edge, such as IoT devices (de Silva et al., 2022)
Additional considerations to take into account (Figure 65.1).
when utilizing this approach include: The application of the proposed methodology within
The VQE and QA algorithms can be implemented a metaverse context has the potential to enhance the
using either a hybrid quantum-classical approach or a security and reliability of IoT networks. To mitigate
purely quantum approach. The selection of the imple- potential security threats targeting IoT devices and
mentation strategy will be contingent upon various ensure their secure communication, the system utilizes
aspects, including the availability of resources and the quantum deep learning techniques to implement trust-
desired performance objectives (Li et al., 2022). based routing (Guan and Morris, 2022).
A diverse array of machine learning methodologies The subsequent expression is a trust-oriented rout-
can be employed for the training of VQE and quan- ing mechanism designed for IoT within the metaverse,
tum approximate optimization algorithms. The selec- leveraging quantum deep learning techniques such as
tion of the right machine learning method will depend VQE and QA.
on the specific characteristics and requirements of the
problem under consideration. R = f(VQE(QA(x)), T)
There are multiple methodologies available for the
implementation of VQE and quantum approximate where:
512 Quantum deep learning on driven trust-based routing framework for IoT in the metaverse context
should be proportionate to its level of significance or The trust-based routing framework, which is pro-
resource abundance (El Saddik, 2023). pelled by quantum learning (VQE + QA), encounters
Trust scores are employed to ensure a secure and notable challenges:
reliable routing path. In the event that an IoT device
possesses a low trust score or is recognized as suscep- • In order to perform this task, it is important to
tible to attacks, it would be advisable to refrain from have access to a quantum computer.
transmitting data over said device. • The system is susceptible to interference caused
The presence of constraints ensures that each IoT by quantum computer noise.
device is exclusively utilized on a singular route. • Training the VQE and quantum approximate
A framework for quantum trust-based routing optimization algorithm, as well as encoding the
driven by VQE and question-answer (QA)-based deep routing problem into a quantum circuit, might
learning. pose significant challenges.
In order to tackle the optimization difficulty dis-
cussed earlier, we provide a trust-based routing Algorithm
architecture that is powered by quantum deep learn-
ing, namely the combination of VQE and quantum Initialize Quantum Deep Learning Environment:
annealing (QA). Initialize Quantum Circuit for VQE
One can employ a quantum circuit for the purpose Initialize Quantum Circuit for QA
of encoding the routing problem. In order to accom- Function VQE-Based_Trust_Modeling(device_
plish this task, it is possible to represent each IoT interactions_data):
node as a quantum bit (qubit) and each potential con- Use VQE to determine the ground state of device
nection as a trajectory inside the quantum circuit. The interactions
characteristics of a quantum circuit can be adjusted to Model trust levels based on historical and real-
accommodate considerations of trustworthiness and time interactions
relevance of IoT devices. Return trust_levels
The objective is to determine the ground state of a Function QA-Based_Routing_Optimization(trust_
quantum circuit using the VQE algorithm. The rout- levels, current_routing_paths):
ing solution that possesses the minimum energy state Use QA to find the optimal routing path based on
is considered to be optimal. trust_levels
The global minimum of the energy function can be Avoid paths with low trust_levels
determined via quantum annealing (QA). Given that Return optimal_routing_path
the VQE algorithm is limited to identifying local min- Function Quantum_Deep_Learning_
ima, this characteristic becomes crucial. Routing(device_interactions_data,
The advantages of employing a trust-based routing current_routing_paths):
system propelled by quantum learning, namely the trust_levels = VQE-Based_Trust_Modeling(device_
combination of VQE and quantum annealing (QA), interactions_data)
are significant. optimal_path = QA-Based_Routing_
The trust-based routing architecture that utilizes Optimization(trust_levels, current_routing_paths)
the quantum deep learning algorithm (VQE + QA)
presents several advantages when compared to con- If optimal_path is valid:
ventional routing methods: Route data through optimal_path
Else:
• Utilization of this approach proves to be more Flag for manual review or fallback to traditional
efficient in identifying the optimal routing alter- routing
native, particularly within extensive and intricate
network systems. Return routing_status
• Enhancing the safety and reliability of routing On IoT Device Data Request in Metaverse:
can be achieved by considering the trust scores device_interactions_data = Fetch historical and
associated with IoT devices. real-time interactions
• The system exhibits enhanced resilience to varia- current_routing_paths = Fetch available paths for
tions in both traffic volume and network topol- data routing
ogy.
routing_status = Quantum_Deep_Learning_
Challenges encountered in the trust-based routing Routing(device_interactions_data,
framework for quantum deep learning (VQE + QA). current_routing_paths)
514 Quantum deep learning on driven trust-based routing framework for IoT in the metaverse context
Table 65.2 Simulation parameters and quantum parameters.
Quantum parameters
Qubits Number of quantum bits used in the quantum e.g., 4, 8, 16, 32, ...
processor
Variational form The choice of variational form or ansatz in VQE UCCSD, Ry, Rz, RyRz, custom ...
Optimizer Classical optimization algorithm used in VQE COBYLA, L-BFGS-B, SPSA, etc.
Entanglement Specifies how qubits are entangled in the Linear, full, circular, custom
quantum circuit
Max iterations (VQE) The maximum number of iterations for VQE e.g., 100, 500, 1000
convergence
Quantum annealing schedule The annealing time or schedule Linear, quadratic, custom
Annealing time Duration of the annealing process e.g., 10 ms, 20 ms, 50 ms
IoT & metaverse parameters
Number of IoT devices Total number of IoT devices in the metaverse e.g., 100, 500, 1000, 10k
simulation
Connectivity model The model dictating how IoT devices are Random, scale-free, small-world, grid
interconnected
Trust evaluation interval How often the trust value for a route or device e.g., 10 s, 60 s, 5 min
is re-evaluated
Trust threshold Minimum trust value required for a route to be e.g., 0.5, 0.7, 0.9
considered valid
Initial trust value Initial trust assigned to devices/routes e.g., 0.5, 1.0
Trust decay rate Rate at which trust value degrades over time e.g., 0.01, 0.05 per minute
without positive reinforcement
Simulation parameters
Simulation time Total time for which the simulation is run e.g., 1 hour, 24 hours
Data packet generation rate Rate at which data packets are generated by IoT e.g., 1 packet/s, 10 packets/s
devices
Malicious node ratio Percentage of nodes that behave maliciously or e.g., 5%, 10%, 20%
unpredictably
Routing protocol The routing protocol employed (with or Classical, quantum-enhanced
without quantum trust considerations)
Traffic model The pattern or model dictating data traffic Uniform, bursty, cyclic
generation
Experiment Qubits Variational Optimizer IoT Connectivity Trust Average Successful Detected
# form devices model threshold packet routes (%) malicious
latency (ms) nodes
Average packet latency – This metric denotes the integration of IoT devices becomes increasingly
the average duration required for a packet to go embedded within its structure. This quantum deep
from its point of origin to its ultimate destination. learning-driven approach sets the stage for the devel-
A lower numerical value corresponds to improved opment of future routing frameworks that are secure,
performance. efficient, and based on trust. It offers the potential
Successful routes (%) – Taking into consideration for the harmonious coexistence of numerous devices
the trust-based framework, this particular indication and services inside the constantly evolving metaverse,
provides insight into the percentage of data packets thereby ensuring a dynamic and resilient infrastruc-
that successfully reached their intended destination. A ture for digital interactions
bigger share corresponds to increased reliability and
efficiency of the route. References
Detected malicious nodes – The following data rep-
resents the cumulative count of simulated IoT devices Lokes, S., Sakthi Jay Mahenthar, C., Parvatha Kumaran,
that exhibit potentially detrimental or unpredictable S., Sathyaprakash, P., and Jayakumar, V. (2022). Im-
plementation of quantum deep reinforcement learn-
behavior.
ing using variational quantum circuits. 2022 Int.
Conf. Trend. Quan. Comput. Emerg. Busin. Technol.
Conclusion (TQCEBT), 1–4.
Incudini, Massimilianoc, Michele Grossi, Antonio Manda-
The increasing prevalence of IoT in the wide realm rino, Sofia Vallecorsa, Alessandra Di Pierro, and David
of the metaverse presents ongoing challenges to tra- Windridge. (2023). The Quantum Path Kernel: a Gen-
ditional routing and security paradigms. Our inquiry eralized Neural Tangent Kernel for Deep Quantum
has unveiled the revolutionary potential of VQE and Machine Learning. IEEE Transactions on Quantum
quantum annealing (QA) in the context of a quantum Engineering, 4, DOI: 10.1109/TQE.2023.3287736.
deep learning-driven framework, showcasing their Gupta, S., Mohanta, S., Chakraborty, M., and Ghosh, S.
capabilities. (2017). Quantum machine learning-using quantum
The successful modeling of trust by the VQE estab- computation in artificial intelligence and deep neu-
lishes a fundamental level of safety by leveraging his- ral networks: Quantum computation and machine
learning in artificial intelligence. 2017 8th Ann. In-
torical and current device interactions to assess the
dus. Autom. Electromec. Engg. Conf. (IEMECON),
dependability of individual nodes. In order to facili- 268–274.
tate the smooth transmission of data traffic in align- Baliuka, A., Stöcker, M., Auer, M., Freiwang, P., Weinfurter,
ment with the trust models established by VQE, the H., and Knips, L. (2023). Deep learning based TEM-
expertise of QA is used to devise routing paths that PEST attacks on a quantum key distribution sender.
are optimized for efficiency. 2023 Conf. Lasers Electro-Opt. Eur. European Quan.
The integration of quantum algorithms and deep Elec. Conf. (CLEO/Europe-EQEC), 1–1.
learning in this complementary approach offers a Suryotrisongko, H and Musashi, Y. (2021). Hybrid quan-
promising solution to address the significant chal- tum deep learning with differential privacy for botnet
lenges encountered in IoT environment within the DGA detection. 2021 13th Int. Conf. Inform. Comm.
metaverse. The effectiveness of this paradigm is sup- Technol. Sys. (ICTS), 68–72.
Jain, S., Gandhi, A., Singla, S., Garg, L., and Mehla, S.
ported by empirical evidence demonstrating a signifi-
(2022). Quantum machine learning and quantum
cant reduction in the occurrence of security breaches, communication networks: The 2030s and the future.
enhanced efficiency in data transfer, and the reinforce- 2022 Int. Conf. Comput. Model. Simul. Optim. (IC-
ment of network infrastructure. CMSO), 59–66.
Quantum methodologies, exemplified by the one Ren, W., Li, Z., Li, H., Li, Y., Zhang, C., and Fu, X. (2020).
elucidated, will prove highly advantageous as the Application of quantum generative adversarial learn-
digital realms within the metaverse expand and
516 Quantum deep learning on driven trust-based routing framework for IoT in the metaverse context
ing in quantum image processing. 2020 2nd Int. Conf. Guan, J. and Morris, A. (2022). Extended-XRI body inter-
Inform. Technol. Comp. Appl. (ITCA), 467–470. faces for hyper-connected metaverse environments.
Chen, S. Y.-C., Huck Yang, C.-H., Qi, J., Chen, P.-Y., Ma, 2022 IEEE Games Entertain. Media Conf. (GEM),
X., and Goan, H.-S. (2020). Variational quantum cir- 1–6.
cuits for deep reinforcement learning. IEEE Acc., 8, Bujari, A., Calvio, A., Garbugli, A., and Bellavista, P. (2023).
141007–141024. A layered architecture enabling metaverse applica-
Liu, G.-X., Liu, J.-F., Zhou, W.-J., and Wu, L. (2022). In- tions in smart manufacturing environments. 2023
verse design local-density-of-states via deep learning IEEE Int. Conf. Metav. Comput. Netw. Appl. (Meta-
in quantum nanophotonics. 2022 Asia Comm. Pho- Com), 585–592.
ton. Conf. (ACP), 2157–2160. Gera, T., Singh, J., Mehbodniya, A., Webber, J. L., Shabaz,
Gupta, B. B., Gaurav, A., Chui, K. T., Wang, L., Arya, L., M., and Thakur, D. (2021). Dominant feature selec-
Shukla, A., and Peraković, D. (2023). DDoS attack de- tion and machine learning-based hybrid approach
tection through digital twin technique in metaverse. to analyze android ransomware. Sec. Comm. Netw.,
2023 IEEE Int. Conf. Cons. Elec. (ICCE), 1–5. 2021, 1–22.
Li, K., Cui, Y., Li, W., Lv, T., Yuan, X., Li, S., Ni, W., Sim- Bouachir, O., Aloqaily, M., Karray, F., and Elsaddik, A.
sek, M., and Dressler, F. (2022). When internet of (2022). AI-based blockchain for the metaverse: Ap-
things meets metaverse: Convergence of physical proaches and challenges. 2022 Fourth Int. Conf.
and cyber worlds. IEEE Internet of Things J., 10(5), Blockchain Comput. Appl. (BCCA), 231–236.
4148–4173. Brik, B., Moustafa, H., Zhang, Y., Lakas, A., and Subrama-
de Silva, R., Zaslavsky, A., Loke, S. W., Jayaraman, P. P., nian, S. (2023). Guest editorial: Multi-access network-
Abken, A., and Medvedev, A. (2022). A scenario based ing for extended reality and metaverse. IEEE Internet
approach for context query generation. 2022 IEEE of Things Mag., 6(1), 12–13.
Smartworld, Ubiquit. Intel. Comput. Scal. Comput. El Saddik, A. (2023). Keynote speaker: The metaverse: AI-
Comm. Dig. Twin Priv. Comput. Metav. Auton. Trust. powered universe of persistent digital twins. 2023
Veh. (SmartWorld/UIC/ScalCom/DigitalTwin/Pri- IEEE 6th Int. Conf. Multimed. Inform. Proc. Retriev.
Comp/Meta), 1469–1476. (MIPR), xix–xix.
66 Advancing network security paradigms integrating
quantum computing models for enhanced protections
Anand Singh Rajawat1, S. B. Goyal2,a, Chaman Verma3 and Jaiteg Singh4
1
School of Computer Science & Engineering, Sandip University, Nashik, Maharashtra, India
2
City University, Petaling Jaya, 46100, Malaysia
Faculty of Informatics, Department of Media and Educational Informatics, Eötvös Loránd University, 1053 Budapest,
3
Hungary
4
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
This study investigates the incorporation of quantum computing models into established network security paradigms in
order to bolster defense mechanisms against emerging cyber threats. The emergence of quantum computing poses a poten-
tial threat to the effectiveness of conventional cryptographic techniques, hence demanding a fundamental transformation
in the field of network security. Our study focuses on exploring the use of quantum algorithms and quantum key distribu-
tion (QKD) mechanisms to enhance encryption and ensure secure communications. The review commences by providing a
comprehensive examination of the fundamental concepts of quantum computing and its potential ramifications for the field
of cybersecurity. Next, we proceed to explore particular quantum algorithms that provide resilient encryption solutions,
surpassing classical equivalents in terms of both security and efficiency. This study delves deeper into the obstacles encoun-
tered during the introduction of quantum technologies and the potential ramifications they may have on network security
infrastructure. By conducting simulations and theoretical evaluations, we provide evidence to support the effectiveness of
quantum-enhanced security models in preventing advanced cyber-attacks. The results of our study indicate that the imple-
mentation of quantum computing models is not only viable but also essential for the progression of network security in the
contemporary digital age
Keywords: Quantum computing, network security, quantum algorithms, quantum key distribution (QKD), cybersecurity,
cryptographic methods
Introduction standards RSA and ECC, which now serve as the foun-
dation for encryption protocols, face the possibility
In the current epoch characterized by the centrality
of becoming obsolete due to the advent of quantum
of digital information in global communication and
computers. These advanced computing systems pos-
business, the significance of network security has
sess the capability to decrypt RSA and ECC encryp-
reached unprecedented levels. The expeditious evolu-
tions far faster than their classical counterparts. With
tion of cyber threats calls for a fundamental change
the recognition of the imminent threat, the objective
in our approach towards network security. The emer-
of this study is to examine and suggest approaches
gence of quantum computing has introduced possible
for the incorporation of quantum computing models
flaws to traditional encryption systems that were pre-
inside network security frameworks. The objective is
viously considered impervious. This study explores
to not alone mitigate the risks brought forth by quan-
the incorporation of quantum computing models into
tum computing, but also to leverage its capabilities
the domain of network security, potentially leading
in order to strengthen network defenses. This paper
to a transformative impact on data and communica-
investigates the concept of quantum key distribution
tion protection within the digital sphere. The advent
(QKD), a cryptographic protocol (Ceylan and Yılmaz,
of quantum computing marks the dawn of a novel era
2021) that leverages principles from quantum physics
in computational capabilities. In contrast to conven-
to establish a secure communication channel, hence
tional computers that operates on binary bits (0s and
ensuring resistance against interception or eavesdrop-
1s), quantum computers (Mukherjee and Kumar Barik,
ping. Additionally, our research encompasses the field
2020) employ quantum bits, or qubits, which enable
of post-quantum cryptography, which entails the
the representation and processing of intricate datas-
creation and refinement of cryptographic algorithms
ets with more efficiency. The substantial advancement
that offer robust security against both quantum and
in computer capacity presents a notable challenge
classical computers. This focus ensures a smooth and
to traditional encryption techniques. The encryption
a
drsbgoyal@[Link]
518 Advancing network security paradigms integrating quantum computing models
uninterrupted transition in the face of the increasing in order to safeguard quantum networks against
prevalence of quantum computing. The incorpora- advanced cyber threats. The research presents a dis-
tion of quantum computing into the realm of net- tinctive methodology for analyzing vulnerabilities in
work security presents inherent difficulties. Careful quantum internet, so offering a useful contribution
consideration is required for the issues pertaining to towards the advancement of secure quantum commu-
scalability, interoperability with existing infrastruc- nication systems.
ture, and the current early stage of development in The study did by Kato et al., (2021) introduces an
quantum technology. The objective of this paper is innovative approach to quantum network coding,
to tackle the aforementioned difficulties by provid- which aims to improve the security and efficiency of
ing valuable insights on strategies to overcome them, quantum networks. The researchers have devised a
ultimately leading to the attainment of a more robust quantum network coding protocol that ensures secu-
and secure digital landscape. As we find ourselves on rity in a single-shot manner. This protocol is specifi-
the cusp of a quantum revolution, it becomes crucial cally designed to be very efficient for multiple unicast
to reconsider and reorganize our existing network networks, which are commonly seen in quantum
security paradigms. This research aims to provide a communication systems.
scholarly contribution to the ongoing discussion by Diamanti’s (2019) research showcases the tangible
presenting a comprehensive plan for incorporating benefits associated with the use of photonic systems in
quantum computing models into the field of network the field of quantum computing, particularly in terms
security. This integration is expected to enhance the of bolstering security measures and improving oper-
security and resilience of digital infrastructure, partic- ational efficiency. This study presents a comprehen-
ularly in light of the ever-evolving landscape of cyber sive examination of the practical implementations of
threats. quantum technology in the realm of network security.
It emphasizes the discernible advantages and progres-
Related work sions that quantum systems can provide in compari-
son to classical systems.
In recent years, there have been notable breakthroughs Table 66.1 presents a concise overview of signifi-
in the science of quantum computing and its utiliza- cant research conducted in the domain of quantum
tion in network security. The subsequent scholarly computing, specifically focusing on its implications
articles offer significant perspectives on diverse facets for network security. Every study makes a substantial
of quantum network security, encompassing the eval- contribution to the progress of quantum computing
uation of security measures and the implementation in this field, while also emphasizing the necessity for
of these measures in real-world industrial settings. additional research, specifically in terms of practical
The paper presented by Zhou et al. (2022) thor- implementations and wider applications.
oughly evaluates the security aspects pertaining to
quantum networks. This study explores the essential Methodology
techniques employed in managing cryptographic keys,
which play a pivotal role in preserving the security An investigation into the fundamentals of quantum
and reliability of quantum communication networks. computing is as follows:
The research conducted by the author centers around Quantum principles: The essay by Zhou et al. (2022)
the examination of vulnerabilities and threat models aims to provide an introduction to the fundamental
that are unique to quantum networks. Their primary principles of quantum computing, with a particu-
objective is to establish a comprehensive framework lar emphasis on features that are highly pertinent
for the assessment and improvement of security mea- to the field of security. The discussion will primar-
sures in these networks. ily revolve on two key concepts: superposition and
The study by Ahmad et al. (2021) investigates the entanglement.
utilization of quantum computing for the purpose of
augmenting security measures inside the Industrial Quantum principles
Internet of Things (IIoT) domain. This paper aims The purpose of this paper is to provide an introduc-
to discuss the escalating apprehensions around the tory overview of the field of quantum computing.
security of networked devices inside industrial envi- Quantum computing is a rapidly evolving area of
ronments. The authors suggest employing quantum research that explores the principles and applications
computing to enhance the security of these networks. of quantum mechanics in the context of information
The study by Satoh et al. (2021) examines potential processing.
vulnerabilities and threats targeting the infrastructure Definition and overview: Quantum computing is a
of quantum internet. The statement underscores the computational paradigm that leverages quantum-
importance of implementing strong security measures mechanical principles, such as superposition and
Applied Data Science and Smart Systems 519
Table 66.1 Comparative table.
Li, et al. (2021) Creating an OSI-like Structures a scalable High implementation Implementation and
quantum Internet quantum Internet for complexity and integration with
paradigm security and efficiency resource needs current infrastructures
require more
investigation
Moreolo, et al. Planning efficient Increases optical Limited to optical Applying to networks
2023) quantum-secure network security with networks; scalability other than optical ones
optical network quantum methods concerns across
communications networks
Aji, et al. 2021 Examining QKD Helps comprehend Concentrates on Research on real-world
network simulation and construct QKD simulation, not implementation issues
systems simulations by application needed
providing an overview
Tong, et al., 2022 Big data research on Combines quantum Limitations in route Route and line model
quantum secure route security with large data and line model picture photo encryption
models and image to improve encryption encryption limitations
encryption
Das and Kule, A new quantum Increases quantum Needs neural network Exploration of simpler,
2022 cryptography error cryptography error integration, which may more efficient quantum
correction method correction reliability be complicated cryptography error
employing artificial correction algorithms
neural networks
between quantum computers and traditional comput- the identification of possible security concerns with
ers enables quantum computers to effectively handle enhanced precision compared to traditional systems.
intricate information with significantly enhanced effi- The field of secure multi-party computation (SMPC)
ciency. The potential of quantum computing to revo- also shows potential for advancements through the
lutionize cryptographic systems renders it a highly utilization of quantum computing. Secure multi-
significant advantage in the realm of network security. party computation (SMPC) (A New Error Correction
Classical cryptography (Li et al., 2021; Brar et al. 2022), Technique in Quantum Cryptography Using Artificial
which encompasses prominent techniques such as Neural Networks 2022) enables the collaborative com-
RSA and ECC, is predicated upon the computational putation of a function by many parties, ensuring the
complexity associated with factoring big numbers or privacy of their respective inputs. Although classical
solving discrete logarithm problems. The aforemen- solutions are characterized by high computational
tioned systems, albeit resistant to conventional com- intensity and inefficiency, quantum computing has
puter attacks, may be susceptible to compromise by the potential to perform these operations with greater
a quantum computer employing Shor’s algorithm. speed and security, hence ensuring strong security
The aforementioned technique exhibits significantly for collaborative computational tasks. Finally, the
improved efficiency in factoring huge numbers when utilization of quantum computing has the potential
compared to the most advanced algorithms now to contribute to the advancement of artificial intel-
available on classical computers. Consequently, this ligence (AI) and machine learning models specifically
advancement poses a significant threat to the secu- designed for enhancing network security. Quantum-
rity of existing cryptographic systems. Therefore, the enhanced machine learning algorithms provide supe-
advancement of quantum computing (Moreolo et al., rior accuracy and efficiency in pattern analysis and
2023) requires the creation of novel cryptographic identification of possible security concerns compared
protocols, sometimes referred to as post-quantum to classical techniques. The possession of this skill has
cryptography, that possess the ability to withstand the potential to play a vital role in mitigating the esca-
quantum attacks. Quantum computing exhibits nota- lating complexity of cyberattacks. The incorporation
ble benefits in the realm of secure communications. of quantum computing models into network security
Quantum key distribution is a cryptographic proto- paradigms presents significant benefits in compari-
col that leverages the fundamental principles of quan- son to traditional computing. Quantum computing is
tum physics to establish a highly secure and robust poised to significantly contribute to the advancement
communication channel. In contrast to conventional of network security, encompassing several aspects
key distribution techniques, which are susceptible to such as reinforcing cryptographic systems against
interception and decryption through the utilization quantum attacks, enhancing secure communications,
of significant processing resources, QKD (Aji et al., and improving anomaly detection. As the advance-
2021) leverages the inherent quantum characteristics ment of this technology progresses, it becomes cru-
of particles such as photons to identify any potential cial for organizations to adjust and make necessary
eavesdropping activities throughout the transmission arrangements for the quantum age, in order to safe-
process. When an individual attempts to observe the guard their networks from emerging cyber risks.
quantum states of these particles without authoriza-
tion, the state of the particle undergoes a transforma- Quantum-enhanced security models
tion as a result of the no-cloning theorem in quantum Quantum key distribution (QKD): This paper delves
mechanics. This alteration serves as a signal to the into the practical application of QKD as a means to
communicating entities, indicating the existence of establish unbreakable encryption, hence guaranteeing
an unauthorized third party. Quantum key distri- the security of communication lines.
bution is rendered fundamentally secure against all Quantum key distribution is an advanced technique
forms of computational advancements, including the within the field of cybersecurity that gives a promising
advent of quantum computing. In addition, the uti- solution for encryption, with the potential to provide
lization of quantum computing has the potential to an invulnerable method of securing data. The incor-
greatly augment network security by enabling more poration of this technology into evolving network
effective anomaly detection mechanisms. Classical security frameworks, especially when integrated with
computers have challenges when confronted with quantum computing models, signifies the arrival of a
the immense quantities (Tong et al., 2022) of data novel era characterized by fortified defense mecha-
and intricate patterns that are essential for achiev- nisms against progressively intricate cyber risks. The
ing proficient anomaly detection in extensive net- fundamental basis of QKD is rooted in the concepts
works. Quantum computers provide the capability to of quantum physics, with a primary focus on utiliz-
do intricate computations and concurrently analyze ing the characteristics of photons to enable secure
extensive datasets, hence enabling them to expedite communication. The fundamental principle at the
Applied Data Science and Smart Systems 521
center of this discussion is the Heisenberg Uncertainty scale. An additional aspect of interest pertains to
Principle, which posits that the process of observing the incorporation of QKD systems into pre-existing
a quantum system inherently modifies its state. The network infrastructures. This encompasses both the
aforementioned principle holds significant impor- physical layer of networks and the establishment of
tance in the context of QKD (Wang et al., 2020), as novel protocols and standards capable of accommo-
it signifies that any endeavor to intercept confiden- dating the distinct demands of quantum key distribu-
tial information can be identified, as it will inevita- tion. The establishment of a safe and interoperable
bly alter the state of the observed quantum system. framework for quantum communication necessi-
The BB84 protocol, which was devised by Charles tates the essential involvement of industry, academia,
Bennett and Gilles Brassard in 1984, is widely recog- and government organizations through collabora-
nized as one of the most frequently employed QKD tive efforts. The incorporation of QKD (Djordjevic,
methods. In this experimental procedure, two enti- 2020) into the progression of network security para-
ties, typically denoted as Alice and Bob, engage in digms shows great potential in attaining impregnable
the exchange of photons that exhibit polarization in encryption amongst the ever-evolving landscape of
one of four distinct orientations. The aforementioned cyber threats. Despite the existence of obstacles in
polarizations serve as discrete units of information, achieving widespread adoption, the ongoing progress
specifically denoted as binary digits, encompass- in quantum technologies and the combined endeav-
ing both 0 and 1 values. The fundamental aspect of ors across diverse sectors are gradually surmounting
security within this protocol lies in the utilization of these problems. This progress signifies a noteworthy
two distinct bases for measuring polarizations. These advancement in the pursuit of secure global commu-
bases are selected in a random and independent man- nication networks.
ner by both participating entities. The uncertainty of Quantum algorithms: This study aims to develop
the basis used for measurements by an eavesdropper and analyze quantum algorithms with the potential to
(referred to as Eve) results in any interception and improve threat detection and response systems.
measurement of photons causing a disturbance to
their state. This disturbance serves as an indication Quantum Key Distribution (QKD) - BB84 Protocol
to Alice and Bob, the communicating parties, about
the presence of an eavesdropper. The incorporation Initialization
of QKD into sophisticated network security systems Alice and Bob agree on two sets of basis vec-
assumes paramount importance in the age of quan- tors, say rectilinear (+) and diagonal (×).
tum computing. Traditional encryption techniques,
such as RSA, may have possible vulnerabilities when Key Generation by Alice
confronted with quantum computers. Theoretically, For each bit of the key she wants to send:
these advanced computing systems have the capabil- a. Alice randomly chooses a basis (+ or ×) and
ity to compromise conventional cryptographic sys- a bit value (0 or 1).
tems in considerably less time compared to classical b. She prepares a qubit in the state corre-
computers. Nevertheless, QKD provides a heightened sponding to her choice.
level of security that is theoretically impervious to (e.g., if she chooses + basis and bit 0, she
quantum attacks. This is due to the fact that the secu- prepares the qubit in state |0〉).
rity of QKD is not contingent upon computational c. She sends the qubit to Bob.
complexity, but rather on the basic principles gov-
erning quantum physics. The use of QKD in practi- Measurement by Bob
cal settings poses numerous obstacles. Two primary For each qubit received:
issues in the field of QKD are the practical deploy- a. Bob randomly chooses a basis (+ or ×) to
ment range and the generation and distribution rate measure the qubit.
of cryptographic keys. Conventional QKD systems b. He measures the qubit and records the out-
encounter (Al-Mohammed et al., 2021) a constraint in come (0 or 1).
terms of the maximum distance over which quantum
states may be preserved without experiencing dete- Basis Reconciliation
rioration, often spanning a few hundred kilometers a. After all qubits are transmitted and mea-
when transmitted via fiber-optic cables. Nevertheless, sured, Alice and Bob communicate over a
new technological developments, such as the imple- classical channel.
mentation of quantum repeaters and the utilization b. They reveal to each other which basis they
of satellite-based QKD, are expanding the limits of used for each qubit, but not the bit values.
quantum-secure communication networks, thus facil- c. They discard any bits where they used dif-
itating their potential deployment on a worldwide ferent bases.
522 Advancing network security paradigms integrating quantum computing models
d. The remaining bits form the raw key. // Pseudo-Code for Quantum Algorithm in Threat
Detection using Cryptography
Key Sifting (Optional)
Alice and Bob can further process the raw key QuantumAlgorithm EnhancedThreatDetection:
for errors or eavesdropping detection:
a. They may decide to disclose a portion of // Initialize Quantum Registers
their key to check for discrepancies. Initialize qubits in superposition to represent all
b. If the error rate is acceptable, they proceed; possible data states
otherwise, they abort the protocol. Initialize ancillary qubits for intermediate
calculations
Final Key // Apply Quantum Cryptographic Algorithm
a. The remaining undisclosed part of the raw Function QuantumCryptography():
key, after any necessary error correction Apply Quantum Fourier Transform (QFT) for
and privacy amplification, becomes the data encryption
shared secret key. Use entanglement and superposition for secure
data sharing
The objective of this research is to devise quantum Return encrypted data state
algorithms that can significantly improve threat
detection capabilities. // Quantum Threat Detection Routine
Function
Algorithm design: The development of quantum algo- QuantumThreatDetection(encryptedData):
rithms for threat detection (Djordjevic, I. B. et al. , For each data sample in encryptedData:
2022) entails the formulation of computational pro- Apply Grover’s algorithm to search for
cedures capable of efficiently processing extensive anomalies
datasets, hence enabling the identification of possible If anomaly detected:
threats with enhanced precision compared to conven- Mark the data sample as a potential threat
tional algorithms.
Return list of potential threats pose significant difficulties for unauthorized entities
// Main Execution attempting to decipher the data in the absence of the
Function Execute(): corresponding quantum key or algorithm employed
encryptedData = QuantumCryptography() in the encryption procedure.
potentialThreats = QuantumThreatDetection( In order to obtain the original data, the encrypted
encryptedData) state undergoes the application of the inverse Quantum
Measure and collapse qubits to retrieve poten- Fourier Transform (IQFT) throughout (Mahdi, S. S. et
tial threats al., 2022) the decryption process. The equation for
Return potentialThreats decryption utilizing the Inverse Quantum Fourier
Transform (IQFT) can be expressed as in Equation (2)
// Execute the algorithm
result = [Link]() IQFT(QFT(|y〉)) = |y〉..(2)
Print(result)
Quantum threat detection: The utilization of Grover’s
Initialization: Quantum registers, also known as algorithm, a quantum search technique, enables an
qubits, are initially set in a superposition state, which efficient search through encrypted data in order to
encompasses the representation of all potential data identify potential dangers or anomalies.
states. Execution and measurement: The primary execution
Quantum cryptography: A function is employed to function is responsible for executing the cryptography
implement quantum cryptography techniques, such and threat detection functions. Ultimately, the qubits
as Quantum Fourier Transform (QFT), in order to undergo measurement in order to induce the collapse
encrypt the data. The implementation of this measure of their respective states and subsequently extract per-
guarantees the preservation of data security through- tinent data regarding potential hazards.
out the detection procedure (Geddada and Lakshmi, Quantum machine learning: Quantum machine learn-
2022). ing (QML) algorithms possess the capability to effi-
ciently process and analyze data in manners that are
The integration of Quantum Fourier Transform not practical for traditional computing systems. This
(QFT) into the advancement of network security par- enables the timely identification of advanced cyber
adigms, particularly in the context of incorporating threats, especially those concealed within extensive
quantum computing models for increased protection, datasets.
can be conceptualized by employing QFT in specific
applications of quantum algorithms. An example of In the domain of quantum machine learning
a potential application is within the realm of encryp- (QML) applied to network security, specifically in the
tion and decryption procedures, where the utilization realm of identifying intricate cyber risks within exten-
of QFT serves to modify quantum states in a manner sive datasets, it is vital to examine an equation that
that augments the level of security. The equation that embodies a quantum-empowered machine learning
serves to demonstrate the aforementioned principle is algorithm. The fundamental component of such an
illustrated below. algorithm generally encompasses a quantum adapta-
Let us consider a quantum state |y〉 that serves as tion of a conventional machine learning model, such
a representation for a sequence of data bits within a as a quantum neural network or a quantum decision
quantum encryption scheme. The utilization of quan- tree.
tum field theory (QFT) in the context of encryption The aforementioned equation exemplifies the fun-
can be expressed as in Equation (1). damental concept that the application of the inverse
quantum Fourier transform to a state that has under-
(1) gone the QFT results through the Equation (2) in the
retrieval of the original state |y〉. This observation
serves as evidence for the practicality of implement-
where, in equation (1) N is the number of qubits. | ing secure encryption and decryption inside quantum-
jk〉 are the basis states e2πijk⁄N represents the complex enhanced network security systems.
exponential fact. A simplified depiction of a quantum machine learn-
In the present situation, quantum field theory (QFT) ing algorithm designed for the purpose of threat iden-
is employed to encode the data into a quantum state, tification is as follows:
exhibiting a level of security that surpasses classical
methodologies. The intricate nature and interconnec- y(x) = ∪(q,X)|y0〉...(3)
tion brought about by quantum field theory (QFT)
524 Advancing network security paradigms integrating quantum computing models
The fundamental component of a QML algorithm In Equation (6), the observed increase in performance
is the unitary operation U, which serves as the cen- is noteworthy, particularly when considering larger
tral element of the learning model, akin to the weights values of N. This is particularly relevant in network
and structure found in a classical neural network. The security contexts, where the necessity to scan exten-
optimization of these processes occurs during the sive datasets for threat detection is prevalent. Hence,
training phase with the objective of minimizing a loss the optimization of quantum search algorithms
function. In the domain of network security, this loss assumes significance in augmenting the efficacy and
function is typically associated with the accuracy of promptness of network security systems.
threat detection.
One notable benefit of utilizing QML (Dong, Y., Results
Zhou, et al., 2022) in this particular context is in its
capacity to expedite the processing and analysis of Simulation environment: Quantum computing simu-
extensive and intricate datasets in comparison to con- lations are employed to evaluate the proposed models
ventional algorithms. This is primarily attributed to within a controlled experimental setting.
the utilization of quantum superposition and entan- Speed and efficiency: This study aims to analyze the
glement, which enable enhanced computational capa- enhancements in processing speed and efficiency that
bilities. The acceleration provided by this technology arise from the utilization of quantum algorithms for
facilitates the timely identification of complex threats the purposes of threat detection and response.
that may otherwise go unnoticed or require a signifi- Accuracy in threat detection: This analysis aims
cant amount of time to be detected by conventional to assess the precision of quantum algorithms in
approaches. threat detection when compared to conventional
Optimization of quantum search algorithms – methodologies.
Quantum search algorithms, such as Grover’s algo- Encryption and data protection evaluates the effi-
rithm, has the capability to perform searches on cacy of quantum encryption techniques, such as
unsorted databases at an exponentially accelerated quantum key distribution (QKD), in augmenting the
rate compared to classical algorithms. The adaptation level of data security (Table 66.1).
of these strategies for network security has the poten-
tial to substantially decrease the duration required for Explanation of Table 66.1
threat detection.
In order to elucidate the optimization of quantum QuantumNetSim: This study examines the potential
search algorithms, such as Grover’s algorithm, within application of QML techniques in the detec-
the realm of network security, it is possible to con- tion of threats within virtual private networks
struct an equation that showcases the improvement in (VPNs), emphasizing the evaluation of accuracy
performance when compared to classical techniques. and speed as crucial performance indicators.
One notable example is Grover’s algorithm, which is CyberQ-virtual lab: This study examines the robust-
renowned for its capacity to do searches on unsorted ness of quantum cryptography within corporate
databases in O(N) time complexity, where N is the networks, with a specific emphasis on evaluating
total amount of objects included within the database. the effectiveness of encryption algorithms and the
In contrast, the traditional equivalent necessitates efficiency of quantum key exchange protocols.
O(N) time for doing the identical work. The optimi- Quantum-ready testbed: This study evaluates the fea-
zation equation in the context of network security can sibility of integrating quantum models into In-
be expressed as follows, where Equation (4 and 5) ternet of Things (IoT) networks, with a specific
Tquantum and Tclassical denote the respective durations of focus on examining compatibility and scalability
quantum and classical algorithms: aspects.
Applied Data Science and Smart Systems 525
Table 66.1 Encryption and data protection
Simulation Description Focus area Quantum model Network type Key metrics
environment used tested
Table 66.2 A comparative analysis of the performance of QML algorithms and standard (classical) machine.
Quantum cloud simulator: This study places empha- terprise networks, facilitating the measurement
sis on the evaluation of quantum entanglement- of predicted accuracy and response time.
based protocols within cloud networks, specifi-
cally examining aspects related to data safety A comparative analysis of the performance of QML
and network performance measures. algorithms and standard (classical) machine learn-
AI-Quantum Framework: Integrating artificial intel- ing methods in the detection of sophisticated cyber
ligence (AI) with quantum computing enables threats (Table 66.2).
enhanced threat detection capabilities within en-
526 Advancing network security paradigms integrating quantum computing models
Search speed Linear time complexity Quadratic speedup Quantum algorithms explore unsorted datasets
(Square root of N) exponentially quicker
Data scalability Decreases with large Remains efficient for Quantum algorithms scale with data without
datasets large datasets losing performance
Accuracy High (varies by High and consistent Quantum algorithms consistently detect threats
algorithm)
Resource High for large datasets Lower compared to Quantum computing uses less for similar tasks
utilization classical algorithms
Threat Varies (generally Significantly reduced Faster cyber threat detection due to efficient
detection time slower) search
Adaptability to Moderate High Quantum algorithms can quickly adapt to
new threats advanced cyberattacks
Quantum Not applicable Inherently resistant to Protects against quantum computing dangers
resilience quantum attacks
Implementation Low to moderate High (due to current Quantum algorithm implementation is
complexity stage of quantum tech) complicated and requires expertise
Energy Moderate Higher Quantum computers can perform complex
efficiency calculations with less energy
Applied Data Science and Smart Systems 527
Table 66.5 Measurements of electronic components.
and limitations. By comparing and contrasting these research has examined the potential of quantum com-
algorithms, researchers. puting to revolutionize threat detection and response
mechanisms, highlighting the significance of quan-
Explanation of Table 66.3 tum algorithms as a crucial answer in the continu-
ous fight against more advanced cyber threats. The
Time to detect threat: The utilization of the quantum
field of quantum computing has exhibited remarkable
algorithm yields a substantial decrease in the
computational powers that hold the promise of fun-
duration required to identify potential risks in
damentally transforming our approach to network
network security, hence enhancing its efficacy as
security. The utilization and advancement of quantum
a tool for promptly detecting and responding to
machine learning algorithms have provided insights
attacks in real-time.
into a prospective era where the identification and
Accuracy of threat detection: The accuracy of threat
mitigation of cyber threats can be accomplished with
detection has seen a discernible enhancement, a
unparalleled efficiency and precision. The remarkable
critical factor in the mitigation of false positives
aspect of quantum algorithms is in their capacity to
and false negatives within the realm of network
handle and analyze extensive datasets in manners that
security.
are unattainable for classical computers. The afore-
Data processed per second: The quantum algorithm
mentioned feature not only facilitates the timely iden-
exhibits a notable capacity for processing data at
tification of complex threats, particularly those that
an accelerated pace, indicating its potential for
are concealed within extensive data streams, but also
effectively managing extensive network data, a
amplifies the scalability and agility of network security
prevalent occurrence in contemporary network
systems. Furthermore, the incorporation of quantum
settings.
models into the realm of network security serves not
Energy efficiency: Quantum algorithms include inher-
only to uphold alignment with the progressing intri-
ent computational efficiency and tend to exhibit
cacy of cyber dangers, but also to proactively surpass
greater energy efficiency, so making them a sig-
them. The advent of quantum cryptography, exempli-
nificant factor to consider in the context of sus-
fied as QKD, presents an encryption technique that
tainable computing practices.
is potentially impervious to decryption, hence offer-
ing a degree of security that is presently unachievable
A visual representation of the potential superiority of
through classical cryptographic methodologies. The
quantum algorithms over classical algorithms in the
significance of this improvement is particularly evi-
domains of search and threat detection is illustrated
dent in a contemporary context where conventional
in Table 66.4. This advantage is expected to grow
encryption techniques are progressively susceptible
as quantum computing technology progresses and
to quantum-level risks. Nevertheless, similar to other
becomes more widely available.
nascent technologies, the process of incorporating
quantum computing into network security encoun-
Conclusion ters several obstacles. The domain of quantum com-
The incorporation of quantum computing models puting is currently in its early developmental phase,
into network security paradigms represents a nota- and the practical deployment of this technology on a
ble progression in the domain of cyber defense. This large scale continues to present significant challenges.
528 Advancing network security paradigms integrating quantum computing models
There is a need to solve many challenges pertaining Optic. Netw. (ICTON), 1–4. [Link]
to hardware limits, algorithmic complexity, and the ICTON59386.2023.10207347.
cultivation of a proficient workforce in the field of Aji, A., Jain, K., and Krishnan, P. (2021). A survey of quan-
quantum technology. Furthermore, it is imperative for tum key distribution (QKD) network simulation plat-
forms. 2nd Glob. Conf. Adv. Technol. (GCAT), 1–8.
the cybersecurity community to maintain a state of
[Link]
constant vigilance about the potential malevolent use
Tong, L., Xia, P., and Lv, T. (2022). Research on quan-
of quantum computing. This emphasizes the neces- tum secure route model and line model image en-
sity for ongoing exploration and advancement in the cryption technology based on Big Data technology.
realm of security solutions that are resistant to quan- Int. Conf. Cloud Comput. Big Data Appl. Softw.
tum threats. Engg. (CBASE), 115–118. [Link]
CBASE57816.2022.00028.
G. Das and M. Kule. (2022). A New Error Correction Tech-
References
nique in Quantum Cryptography using Artificial Neu-
P. Mukherjee and R. Kumar Barik. (2022). Fog- ral Networks, 2022 IEEE 19th India Council Interna-
QKD:Towards secure geospatial data sharing mech- tional Conference (INDICON), Kochi, India, 1–5, doi:
anism in geospatial fog computing system based 10.1109/INDICON56171.2022.10040091.
on Quantum Key Distribution, 2022 OITS Inter- Wang, R., Wang, Q., Kanellos, G. T., Nejabati, R., Simeoni-
national Conference on Information Technology dou, D., Tessinari, R. S., Hugues-Salas, E., Bravalheri,
(OCIT), Bhubaneswar, India, 485–490, doi: 10.1109/ A., Uniyal, N., Muqaddas, A. S., Guimaraes, R. S., Di-
OCIT56763.2022.00096. allo, T., and Moazzeni, S. (2020). End-to-end quan-
Ceylan, O. S. and Yılmaz, I. (2021). QDNS: Quantum tum secured inter-domain 5G service orchestration
dynamic network simulator based on event driv- over dynamically switched flex-grid optical networks
ing. Int. Conf. Inform. Sec. Cryptol. (ISCTURKEY), enabled by a q-ROADM. J. Lightwave Technol., 38(1),
2021, 45–50. [Link] 139–149. [Link]
KEY53027.2021.9654307. Brar, P. S., Shah, B., Singh, J., Ali, F., and Kwak, D. (2022).
Zhou, H., Lv, K., Huang, L., and Ma, X. (2022). Quantum Using modified technology acceptance model to evalu-
network: Security assessment and key management. ate the adoption of a proposed IoT-based indoor disas-
IEEE/ACM Trans. Netw. 30(3), 1328–1339. https:// ter management software tool by rescue workers. Sen-
[Link]/10.1109/TNET.2021.3136943. sors, 22(5), 1866. [Link]
Ahmad, S. F., Ferjani, M. Y., and Kasliwal, K. (2021). En- Al-Mohammed, H. A., Al-Ali, A., Yaacoub, E., Abualsaud,
hancing security in the industrial IoT sector using K., and Khattab, T. (2021). Detecting attackers dur-
quantum computing. Cir. Sys. (ICECS), 28th IEEE ing quantum key distribution in IoT networks us-
Int. Conf. Elec., 2021, 1–5. [Link] ing neural networks. IEEE Globecom Workshops
ICECS53924.2021.9665527. (GC Wkshps), 1–6. [Link]
Satoh, T., Nagayama, S., Suzuki, S., Matsuo, T., Hajdušek, shps52748.2021.9681988.
M., and Meter, R. V. (2021). Attacking the quantum Djordjevic, I. B. (2020). Secure, global quantum commu-
internet. IEEE Trans. Quan. Engg., 2, 1–17. https:// nications networks. 22nd Int. Conf. Trans. Optic.
[Link]/10.1109/TQE.2021.3094983. Netw. (ICTON), 1–5. [Link]
Kato, G., Owari, M., and Hayashi, M. (2021). Single-shot TON51198.2020.9203116.
secure quantum network coding for general multiple Geddada, V. J. and Lakshmi, P. V. (2022). Distance based
unicast network with free one-way public communica- security using quantum entanglement: A sur-
tion. IEEE Trans. Inform. Theory, 67(7), 4564–4587. vey. 13th Int. Conf. Comput. Comm. Netw. Tech-
[Link] nol. (ICCCNT), 1–4. [Link]
Diamanti, E. (2019). Demonstrating quantum advantage ICCCNT54827.2022.9984468.
in security and efficiency with practical photonic sys- Mahdi, S. S. and Abdullah, A. A. (2022). Improved se-
tems. 21st Int. Conf. Trans. Optic. Netw. (ICTON), curity of SDN based on hybrid quantum key dis-
1–2. [Link] tribution protocol. Int. Conf. Comp. Sci. Softw.
Li, Z., Xue, K., Li, J., Yu, N., Liu, J., Wei, D. S. L., Sun, Engg. (CSASE), 36–40. [Link]
Q., and Lu, J. (2021). Building a large-scale and wide- CSASE51777.2022.9759635.
area quantum Internet based on an OSI-alike model. 19. Dong, Y., Zhou, Y., and Yao, Q. (2022). Character-
China Comm. 18(10), 1–14. [Link] ization of nonlocality in chained quantum networks.
JCC.2021.10.001. IEEE 22nd Int. Conf. Softw. Qual. Reliab. Sec. Com-
Moreolo, M. S., Iqbal, M., Nadal, L., and Muñoz, R. (2023). pan. (QRS-C), 515–520. [Link]
Efficient solutions for quantum secure communica- QRS-C57518.2022.00082.
tions in future optical networks. 23rd Int. Conf. Trans.
67 Optimizing 5G and beyond networks: A comprehensive
study of fog, grid, soft, and scalable computing models
S. B. Goyal1,a, Anand Singh Rajawat2, Jaiteg Singh3 and Tony Jan4
1
City University, Petaling Jaya, 46100, Malaysia
2
School of Computer Science & Engineering, Sandip University, Nashik, Maharashtra, India
3
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
4
Department of Information Technology, Torrens University, Australia
Abstract
The emergence of 5G technology and the excitement around future networks beyond 5G have marked the onset of a novel
phase in digital connectivity. This phase is distinguished by exceptionally fast speeds, low delays in data transmission, and
the ability to connect a vast number of devices simultaneously. In order to maximize the capabilities of advanced networks,
it is imperative to thoroughly investigate and include various computing models that can effectively complement and aug-
ment their potential. This research study examines the investigation of four significant computing paradigms such as fog,
grid, soft, and scalable computing and their incorporation into 5G and subsequent networks. Fog computing facilitates the
proximity of computing resources to end-users, resulting in a reduction of latency and an improvement in data processing
speeds. Grid computing is a technique to distributed computing that facilitates the sharing of abundant resources across
networks. This is particularly crucial for effectively managing the substantial volume of data produced by 5G networks. Soft
computing, characterized by its emphasis on artificial intelligence (AI) and machine learning (ML), offers the essential flex-
ibility and decision-making capabilities needed in dynamic network environments. Scalable computing plays a crucial role
in facilitating the efficient scaling of network infrastructures in response to varying demands, which is an essential necessity
for accommodating the dynamic workloads commonly observed in contemporary networks. The objective of this study is
to gain a holistic comprehension of the integration of these models inside 5G and forthcoming network architectures. The
integration being discussed holds the potential to improve network efficiency, stability, and scalability, hence opening up new
possibilities for advancements in network performance and user experience.
Keywords: Computing, grid computing, soft computing, scalable network architectures, 5G technology enhancements, next-
generation network models
a
drsbgoyal@[Link]
530 Optimizing 5G and beyond networks
exhibited by this candidate render it very suitable (AI) and machine learning (ML), with the aim of
for enhancing resource allocation, network manage- addressing intricate challenges within 5G networks.
ment, and decision-making processes within intricate This research enhances the comprehension of how
5G networks. Scalable computing, which is essential soft computing techniques can be utilized to enhance
for addressing the dynamic requirements of contem- network performance and user experience in the con-
porary networks, guarantees the effective expan- text of 5G technology.
sion or contraction of the computing infrastructure Gupta and Singh (2022) in their study provided a
in accordance with varying network loads and user novel approach that integrates a deep reinforcement
demands. The capacity to adapt is of utmost impor- learning framework with a hybrid grey wolf and mod-
tance in ensuring optimal performance and service ified moth flame optimization technique. The objec-
quality within the contexts of 5G and B5G (Meng tive of this approach is to increase load balancing in
et al., 2020) settings. The objective of this article is fog-IoT situations. The significance of this research
to analyses the roles of different computing models, lies in its pioneering methodology for tackling the
evaluate their possible implications, and anticipate complexities associated with load balancing in intri-
how their integration can advance 5G and future cate fog-IoT networks. It presents a unique solution
networks in realizing their complete revolutionary that combines advanced optimization techniques with
potential. Through the comprehensive analysis of deep learning.
these models, our aim is to establish a fundamental
basis for forthcoming advancements and pragmatic Methodology
applications that will significantly influence the tra-
jectory of telecommunications. The swift progression of wireless networks, notably
The paper is organized in the following section such with the introduction of 5G (Khattar et al. 2020,
as the related work, proposed methodology, results Akram et al. 2021) and the expectation of B5G tech-
analysis, discussion, and finally conclusion and future nologies, calls for a reassessment of computing para-
work. digms to effectively accommodate these advanced
networks. The efficacy of traditional cloud-centric
models is progressively diminishing as a result of
Related work
issues related to latency, bandwidth, and processing.
Singh and Kumar (2023) in their paper introduces a This paper provides an examination of many alter-
methodology for maintaining privacy in the aggre- native computing paradigms, namely fog computing,
gation of multidimensional data within smart grid grid computing, soft computing, and scalable comput-
systems, with a particular focus on ensuring secure ing, and their possible implications for the advance-
processing of queries. The proposed framework meets ment of 5G and future networks (Table 67.1).
the crucial requirement for privacy in smart grids by
presenting a comprehensive approach that guarantees Fog computing
data confidentiality, while also facilitating rapid data Fog computing is an extension of cloud computing
aggregation and query processing. The significance that aims to deliver processing, storage, and network-
of the scheme in the era of intelligent infrastructure ing services in closer proximity to data sources and
underscores its relevance in the management and end-users by moving them to the edge of the net-
security of intricate data systems. work. The close proximity between devices results in
Chen et al. (2021) in their work investigate the decreased latency, which is an essential factor for the
optimization of fog radio access networks (F-RAN) in successful operation of 5G applications such as IoT,
the context of 5G technology, by utilizing user mobil- autonomous vehicles, and real-time analytics. The
ity and traffic statistics. The study presents a meth- decentralized architecture of fog computing offers
odology for improving network efficiency and user improved data privacy and security by enabling local
experience in 5G networks by incorporating real-time data processing instead of relying on transmission to
data analytics into F-RAN. The research concentrates a central cloud infrastructure. Nevertheless, the man-
on the utilization of user mobility and traffic data to agement and security of a large quantity of fog nodes
optimize networks, offering vital insights into effi- present considerable obstacles (Khan et al., 2017).
cient resource allocation and network management
strategies in the context of 5G. Grid computing
Divakaran et al. (2022) conducted a technical Grid computing is a process that entails the amal-
investigation on the utilization of soft computing gamation of distributed computing resources, which
techniques to augment the functionality of 5G net- are frequently located in different geographical loca-
works. The study explores a range of soft computing tions, for the purpose of collectively executing a sin-
methodologies, encompassing artificial intelligence gular operation. This approach is very well-suited for
Applied Data Science and Smart Systems 531
Table 67.1 Optimizing 5G and beyond networks.
Model type Key parameters Equations & functions Performance metrics Use cases in 5G
Fog Latency, bandwidth, Data processing time, Response time, data Edge analytics, real-
computing node capacity network latency, resource throughput, energy time IoT applications
allocation efficiency
Grid Computational power, Load balancing algorithm, Computational High-performance
computing data storage, task data transfer rates, resource efficiency, scalability, computing, large-scale
scheduling utilization reliability simulations
Soft Fuzzy logic parameters, Machine learning Accuracy, flexibility, AI-driven network
computing neural network layers, algorithms, optimization robustness management,
genetic algorithm techniques, adaptive control predictive
variables systems maintenance
Scalable Scalability metrics, Dynamic resource scaling, System scalability, Cloud services,
computing resource allocation, virtual network function, resource optimization, dynamic network
virtualization techniques auto-scaling algorithms cost efficiency configurations
intricate computations that need substantial resources, demands. Scalable computing is a critical aspect of
which could be facilitated by 5G networks. Examples network infrastructure that guarantees the ability to
of such computations include large-scale simulations efficiently process and manage substantial volumes of
and data processing. Grid computing is a method that data originating from a multitude of devices, while
enables the use of underutilized resources distributed maintaining optimal performance levels. The key
throughout a network, hence providing a solution obstacle is in the development of systems that pos-
that is cost-effective. Nevertheless, the coordination sess the ability to flexibly adjust to evolving demands
and management of these heterogeneous resources, as while simultaneously upholding efficiency and secu-
well as the guarantee of dependable and continuous rity (Abdali et al., 2021).
service, can present significant challenges.
Integration in 5G and beyond networks
Soft computing The incorporation of these computing paradigms into
Soft computing, in contrast to conventional com- 5G and subsequent generations of networks presents
puting, is concerned with the utilization of approxi- a multitude of advantages.
mation models and the acceptance of imprecision, Enhanced performance: The efficient management
uncertainty, and partial truth in order to attain of high bandwidth and low latency needs in 5G appli-
tractability, robustness, and cost-effectiveness in cations can be achieved by dispersing computing
problem-solving. Soft computing techniques have duties across fog, grid, and scalable systems within
the potential to improve decision-making processes, networks.
adaptability, and learning skills in 5G networks.
These capabilities are essential for effectively man- Proposed Algorithm 1:
aging the complexities of dynamic network envi-
Algorithm steps:
ronments and diverse data sources. The utilization
• Initialization:
of this technology has the potential to enhance the
• Identify the set of computing tasks
efficiency of network traffic, allocation of resources,
T={t1,t2,...,tn} for the 5G network.
and resilience to faults. The primary difficulty is in
• Define the computing resources avail-
the seamless integration of these soft computing
able in fog, grid, and scalable systems:
approaches within established network structures
F={f1,f2,...,fa},
and protocols.
• G={g1,g2,...,gb}, and
• S={s1,s2,...,sc} respectively.
Scalable computing
• Task characterization:
The field of scalable computing is concerned with the
• For each task ti, determine its bandwidth
development of systems that possess the ability to
and latency requirements.
effectively adjust their capacity in response to fluctua-
• Categorize tasks into low-latency (LL),
tions in demand. In the context of 5G (BahraniPour
high-bandwidth (HB), and balanced re-
et al., 2023) and subsequent networks, the ability to
quirements (BR).
scale is of utmost importance because to the dynamic
• Resource assessment:
nature of device quantities and diverse bandwidth
532 Optimizing 5G and beyond networks
Fog Effective at edge-based data Improves network-edge real-time Very massive network topologies
computing processing, lowering latency and data processing for IoT and real- may have scalability concerns
response times time applications
Grid Scalability issues may arise with Ideal for data-intensive Complex management and
computing huge network topologies applications, improving 5G coordination may cause
network processing for complex inefficiencies in dynamic
operations network situations
Soft Adaptable and imprecision- Enhances 5G network decision- May not always deliver
computing tolerant, it handles ambiguous or making, especially in AI and ML appropriate answers, causing
noisy data well performance fluctuation
Scalable Ability to dynamically alter Supports 5G network scalability Problems maintaining
computing resources to demand to meet growing traffic and user performance and efficiency
demands during rapid scaling
connectivity. The primary focus pertains to the pro- to end-users. In the realm of 5G technology, fog com-
vision, administration, and enhancement of these puting’s close proximity to end-users plays a crucial
networks to effectively accommodate an increasingly role in minimizing latency, which is of utmost impor-
vast volume of data and a diverse range of devices and tance for applications that necessitate real-time pro-
services. This discourse explores four crucial comput- cessing. This is particularly relevant for IoT devices,
ing models – fog, grid, soft, and scalable computing autonomous vehicles, and augmented reality, where
that play a significant role in augmenting 5G (Chen et timely data processing is critical.
al., 2019) and subsequent networks.
Latency reduction: Fog computing significantly
Fog computing: Enhancing edge capabilities reduces the duration required for data process-
Fog computing is a paradigm that expands upon the ing and decision-making by conducting these tasks
principles of cloud computing by facilitating the prox- locally instead of transmitting the data to a central-
imity of processing, storage, and networking services ized cloud.
Applied Data Science and Smart Systems 535
Bandwidth optimization: Additionally, it mitigates hence posing a potential constraint on their practical
the limitations of bandwidth by reducing the amount implementation.
of data that must be transmitted over extended dis- Scalable computing: Addressing the needs of expand-
tances. This phenomenon proves to be especially ing networks
advantageous in densely populated urban regions The concept of scalable computing pertains to the
characterized by high levels of network congestion. capacity of a computing system to effectively man-
Nevertheless, fog computing presents several issues age increasing workloads or to be easily expanded
in the areas of security and data management due to in order to accommodate such expansion. Scalability
the decentralized nature of data distribution over mul- holds significant importance in the context of 5G and
tiple nodes. The complexity of maintaining consistent subsequent generations of networks.
security standards and effective data synchronization
Handling growing data volumes: With the increasing
increases in decentralized architectures.
proliferation of connected devices and the exponen-
Grid computing: The concept of resource sharing and tial growth in data generation, the concept of scal-
collaborative processing refers to the practice of pool- able computing has emerged as a means to enhance
ing and utilizing shared resources and engaging in network capabilities while maintaining optimal
cooperative efforts to process tasks or solve problems. performance.
Grid computing refers to the utilization of several
Flexible infrastructure: Scalable computing facili-
distributed computing resources that are intercon-
tates enhanced network flexibility, enabling seamless
nected and typically located in different geographical
adaptation to dynamic demands without necessitat-
regions, with the purpose of collectively executing a
ing a comprehensive restructuring of the underlying
singular operation. The use of a collaborative method
infrastructure.
has the potential to greatly augment the computing
capabilities of 5G networks.
The primary difficulty associated with scalable
Resource optimization: Grid computing enables the computing pertains to the development of systems
consolidation of resources, hence facilitating the effi- that can successfully and economically expand in both
cient management of extensive computations and upward and downward directions, while avoiding the
voluminous datasets. issues of resource underutilization and bottlenecks.
Collaborative processing: It facilitates cooperative The incorporation of fog, grid, soft, and scalable
research and development endeavors, allowing for the computing models into 5G and future networks
smooth collaboration of multiple organizations. offers a comprehensive strategy for augmenting net-
One significant limitation associated with grid com- work capabilities. Each model effectively tackles dis-
puting pertains to the intricate nature of its infra- tinct difficulties and contributes distinct value to the
structure, which presents complexities in terms of network architecture. Fog computing facilitates the
coordination and management of the extensive localization of data processing in close proximity to
system. its source, resulting in a reduction in both latency and
Soft computing: The significance of AI and ML. bandwidth consumption. Grid computing utilizes the
Soft computing approaches, including as neural capabilities of dispersed resources to facilitate cooper-
networks, fuzzy logic, and evolutionary algorithms, ative and efficient execution of computationally inten-
have a significant impact on enhancing the intelli- sive tasks. Soft computing is a field that incorporates
gence and adaptability of 5G networks. AI and ML techniques to enhance the intelligence
and adaptability of network management. On the
Predictive analysis and adaptive learning: Artificial
other hand, scalable computing focuses on enabling
intelligence (AI) and ML algorithms have the capa-
networks to expand and adjust in response to evolv-
bility to forecast network congestion and adaptively
ing requirements. Nevertheless, these advantages are
regulate bandwidth allocation, thereby enhancing the
not devoid of their associated difficulties. Addressing
overall efficiency of the network.
security, data management, infrastructure complexity,
Automated optimization: These models have the and resource optimization are crucial challenges that
capability to automate many network administration must be overcome. Furthermore, the successful incor-
operations, hence decreasing the reliance on human poration of these models into pre-existing network
involvement and mitigating the occurrence of errors. architectures necessitates meticulous strategizing and
Nevertheless, the utilization of data for the pur- implementation. As the progression towards increas-
pose of training these models gives rise to appre- ingly sophisticated network technologies unfolds, it
hensions over privacy and the security of data. becomes evident that the integration of various com-
Furthermore, the intricate nature of these algorithms puting models will play a pivotal role in fully harness-
necessitates substantial computational resources, ing the capabilities of 5G and subsequent networks.
536 Optimizing 5G and beyond networks
The integration of these models can result in networks Meng, Y., Naeem, M. A., Almagrabi, A. O., Ali, R., and Kim,
that exhibit enhanced speed and reliability, while also H. S. (2020). Advancing the state of the fog comput-
demonstrating heightened intelligence, efficiency, and ing to enable 5g network technologies. Sensors, 20(6),
the ability to accommodate the escalating require- 1754.
Singh, A. K. and Kumar, J. (2023). A privacy-preserving
ments of an interconnected global environment.
multidimensional data aggregation scheme with se-
cure query processing for smart grid. J. Supercomput.,
Conclusion 79(4), 3750–3770.
Chen, L., Jiang, Z., Yang, D., and Wang, C. (2021). Fog
The investigation of many computing models, includ- radio access network optimization for 5G leveraging
ing fog, grid, soft, and scalable computing, to enhance user mobility and traffic data. J. Netw. Comp. Appl.,
5G and future networks unveils a landscape abundant 191, 103083.
with potential and innovative possibilities. Each model Divakaran, J., Malipatil, S., Zaid, T., Pushpalatha, M., Patil,
possesses distinct characteristics that, when effectively V., Arvind, C., Joby Titus, T., et al. (2022). Technical
utilized, can greatly contribute to the advancement study on 5G using soft computing methods. Scientif.
and optimization of future networks. Fog computing Program., 2022, 1–7.
has proven its efficacy in reducing latency, enhancing Gupta, S. and Singh, N. (2022). Fog-GMFA-DRL: Enhanced
data processing speed, and improving user experience deep reinforcement learning with hybrid grey wolf
and modified moth flame optimization to enhance the
through its decentralized and edge-centric methodol-
load balancing in the fog-IoT environment. Adv. Engg.
ogy, which involves bringing compute closer to the Softw., 174, 103295.
source of data. Rapid decision-making is of utmost Akram, J., Tahir, A., Munawar, H. S., Akram, A., Kouzani,
importance in 5G networks, specifically in real-time A. Z., and Parvez Mahmud, M. A.. (2021). Cloud- and
applications like IoT and autonomous cars. In contrast, fog-integrated smart grid model for efficient resource
grid computing presents a decentralized approach to utilisation. Sensors, 21(23), 7846.
processing capacity, facilitating the efficient utilization Khan, S., Parkinson, S., and Qin, Y. (2017). Fog computing
of a huge network of resources for the execution of security: A review of current applications and security
intricate and extensive computational operations. The solutions. J. Cloud Comput., 6(1), 1–22.
utilization of this technology within 5G networks has BahraniPour, F., Mood, S. E., and Farshi, Md. (2023). En-
the potential to optimize the management of large- ergy-delay aware request scheduling in hybrid cloud
and fog computing using improved multi-objective CS
scale data, thereby bolstering performance in domains
algorithm. Soft Comput., 1–14.
such as intelligent urban environments and sophis- Abdali, T.-A. N., Hassan, R., Aman, A. H. Md., and Nguy-
ticated data analysis. Soft computing is a field that en, Q. N. (2021). Fog computing advancement: Con-
incorporates flexibility and adaptability into comput- cept, architecture, applications, advantages, and open
ing models, which is crucial for effectively handling the issues. IEEE Acc., 9, 75961–75980.
uncertainties and imprecise information commonly Habibi, P., Farhoudi, Md., Kazemian, S., Khorsandi, S., and
seen in real-world situations. The integration of soft Leon-Garcia, A. (2020). Fog computing: A compre-
computing approaches has the potential to enhance hensive architectural survey. IEEE Acc., 8, 69105–
the resilience and capabilities of 5G networks in man- 69133.
aging dynamic and complex settings, hence providing Ahmadzadeh, S., Parr, G., and Zhao, W. (2021). A review on
communication that is more robust and dependable. communication aspects of demand response manage-
ment for future 5G IoT-based smart grids. IEEE Acc.,
Scalable computing plays a crucial role in effectively
9, 77555–77571.
managing the increasing requirements of network Khattar, N., Singh, J., and Sidhu, J. (2020). An energy effi-
users and devices. The objective is to guarantee that cient and adaptive threshold VM consolidation frame-
the network infrastructure can effectively adjust its work for cloud environment. Wirel. Per. Comm., 113,
capacity to accommodate fluctuating demands while 349–367.
maintaining optimal performance and dependability. Yang, M., Ma, H., Wei, S., Zeng, Y., Chen, Y., and Hu, Y.
The ability to scale is a crucial factor for ensuring the (2020). A multi-objective task scheduling method for
sustainable expansion of 5G networks, particularly fog computing in cyber-physical-social services. IEEE
as we progress towards increasingly data-intensive Acc., 8, 65085–65095.
applications and services. Yin, Z., Xu, F., Li, Y., Fan, C., Zhang, F., Han, G., and Bi, Y.
(2022). A multi-objective task scheduling strategy for
intelligent production line based on cloud-fog com-
References puting. Sensors, 22(4), 1555.
Ahvar, E., Ahvar, S., Raza, S. M., Vilchez, J. M. S., and Lee, Chen, S., Wen, H., Wu, J., Lei, W., Hou, W., Liu, W., Xu, A.,
G. M. (2021). Next generation of SDN in cloud-fog and Jiang, Y. (2019). Internet of things based smart
for 5G and beyond-enabled applications: Opportuni- grids supported by intelligent edge computing. IEEE
ties and challenges. Network, 1(1), 28–49. Acc., 7, 74089–74102.
68 Smart protocol design: Integrating quantum computing
models for enhanced efficiency and security
S. B. Goyal1,a, Sugam Sharma2, Anand Singh Rajawat3 and Jaiteg Singh4
1
City University, Petaling Jaya, 46100, Malaysia
2
CSSM Principal Systems Architect, IOWA State University, USA
3
School of Computer Science & Engineering, Sandip University, Nashik, Maharashtra, India
4
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
In the context of the swiftly progressing domain of digital communications, the imperative for resilient and effective proto-
cols has become increasingly crucial. The incorporation of quantum computing models into the design of protocols signifies
a significant and innovative transformation, presenting unparalleled improvements in terms of both effectiveness and safe-
guarding measures. The present study, investigates the profound impact that quantum computing can have on the evolution
of communication protocols. Quantum computing offers a unique methodology for data processing and transmission by
leveraging the principles of quantum mechanics, including superposition and entanglement. This approach effectively tack-
les the constraints imposed by classical computing models. This study explores the advancements in quantum-enhanced
protocols, with a specific focus on their greater efficiency in data management and inherent security benefits, particularly in
the face of complex cyber threats. The paper moreover explores the obstacles and prospective remedies associated with the
integration of quantum models into practical communication systems, encompassing issues such as compatibility with pre-
existing infrastructure and the ability to scale effectively. The results highlight the important impact of quantum computing
on the development of safe and efficient digital communication, representing a crucial advancement towards more sophisti-
cated and robust network infrastructures.
Keywords: Quantum computing, protocol design, network security, efficiency enhancement, cryptographic algorithms, quan-
tum resilience
a
drsbgoyal@[Link]
538 Smart protocol design: Integrating quantum computing models
cybersecurity concerns and the ever-growing demand about their contribution to a particular field. The
for better data processing rates, the incorporation study presented a noteworthy advancement in the
of quantum computing into protocol design is not practical application of quantum networking, effec-
merely novel but critical. Through the utilization of tively tackling crucial obstacles related to the disper-
quantum physics, it becomes possible to surpass the sion of entanglement.
constraints imposed by conventional computing and The findings by Chiti et al. (2022) in his research
initiate a novel epoch characterized by highly efficient emphasized the potential of quantum-drone networks
and secure digital communication networks. This in metropolitan settings to boost computing and net-
study aims to provide a scholarly contribution to the working capabilities. Contribution to the field: This
emerging subject by conducting a thorough examina- research has introduced a novel direction in the study
tion of the prospective uses and advantages of quan- of quantum networking, particularly within the
tum computing in the progression of protocol design framework of urban and mobile environments.
community. This paper is organized as to represent The findings by Shi and Li (2022) created a pro-
the related work, proposed methodology, results anal- tocol which offered a method for calculating the
ysis, and finally conclusion and future work. cardinality of intersections between numerous par-
ties, while ensuring the confidentiality of individual
Related work datasets.
Contribution to field: This study has enhanced the
The techniques presented in the study by Ji et al. functionalities of safe multiparty computations in
(2019) offer a secure means of comparing private quantum environments, emphasizing the promise of
data while preserving its confidentiality, by using the quantum computing in addressing intricate privacy-
distinct characteristics of quantum entanglement. preserving issues.
Contribution to field: The present study made Table 68.1 present study examines the various
a valuable contribution to the domain of quan- applications of quantum computing in network and
tum cryptography by improving the efficacy of pri- communication systems.
vacy-preserving techniques employed in quantum
communications.
Methodology
The objective of the study design did by Li (2022)
was to optimize the allocation of quantum entangle- The quantum computing and protocol design
ment across network connections, hence improving Design principles: When designing protocols, it
the efficiency and reliability of quantum networks. The is imperative to take into account the distinctive
user’s text does not provide any specific information characteristics of quantum computing. This entails
Shi, 2021 Cardinality of quantum Increases multi- Scalability concerns A need for streamlined
multiparty privacy set party privacy; in larger networks; protocols that scale
intersection quantum-resistant implementation with networks
complexity
Cacciapuoti et Quantum teleportation Novel data transmission Entanglement is Explore resource-
al., 2020 for the quantum method; great security resource-intensive efficient quantum
Internet, combining teleportation
entanglement and classical technologies
communications
Doolittle et Variational quantum Non-locality in noisy Optimizing non- Optimizing quantum
al., 2023 optimization of quantum quantum networks is locality is difficult network non-locality
network non-locality addressed, improving and computationally with more efficient
network robustness intensive methods
Shi, 2022 QuNetSim: Quantum Improves quantum Simulations may Unifying simulation
network software network modeling and not portray network with real-world
framework development complexity application
Shi and Li, Auction of anonymous Provides secure and Auction-specific; little Applying this
2022 quantum sealed bids private quantum generalizability technology to
auctions additional secure
communications areas
Applied Data Science and Smart Systems 539
the conceptualization of protocols that can utilize Design considerations for quantum protocols
quantum parallelism and entanglement in order to Scalability: Although quantum computing presents
enhance both efficiency and security. The principles notable benefits, a primary obstacle in protocol design
of design refer to a set of fundamental guidelines is in the assurance of scalability. At present, quantum
that are employed in the creation and execution of systems have constraints in terms of the stability of
visual compositions (Yang et al., 2022). These prin- qubits and the duration of coherence, both of which
ciples serve utilizing quantum computing for enhanc- have implications for their capacity to execute opera-
ing protocol efficiency and security The emergence of tions on a wide scale. The development of protocols
quantum computing has initiated a novel epoch of that can effectively function within these limitations is
technical potentials, specifically within the realm of of utmost importance.
protocol design. In contrast to classical computing, Error correction: Quantum systems are susceptible to
which operates on binary bits with discrete states of errors as a result of quantum decoherence (El-Latif et
0 or 1, quantum computing employs qubits that can al., 2018) and various sources of noise. The inclusion
exist in superposition, allowing for simultaneous exis- of effective mistake correcting techniques is neces-
tence in several states. This paper explores the essen- sary in order to uphold the integrity of the protocols.
tial design principles required for the development of Quantum error correction codes, such as the surface
protocols that leverage the distinctive characteristics code, provide strategies for mitigating faults in quan-
of quantum computing in order to improve efficiency tum systems while preserving the integrity of quan-
and security. Quantum parallelism is a phenomenon tum states.
that emerges from the inherent capability of quantum
Interoperability: In order for quantum protocols to
computers (Song and Chen, 2020) to concurrently
achieve widespread adoption, it is imperative that
handle many inputs. The ability of qubits to express
they demonstrate compatibility with pre-existing clas-
many states simultaneously is the underlying cause of
sical networks. This entails the development of hybrid
this phenomenon. In the field of protocol design, this
systems capable of executing quantum and conven-
feature can be utilized to do intricate computations in
tional algorithms, hence assuring a smooth integra-
significantly less time compared to conventional com-
tion and transition between the two computational
puters. An example of this may be seen in quantum
paradigms.
algorithms, such as Shor’s algorithm which is used for
the purpose of factoring huge numbers. These algo- Resource optimization: At present, there is a limited
rithms showcase the ability of quantum systems to availability and high cost associated with quantum
efficiently do tasks that would need significant pro- resources (Guo et al., 2019), such as qubits. The opti-
cessing resources on classical systems. This idea is mal utilization of these resources in the design of pro-
applicable to network protocols that require rapid tocols is crucial for practical implementations. This
processing and decision-making, as seen in high-fre- entails the optimization of algorithms with the aim
quency trading or real-time data analysis systems. of minimizing the quantity of qubits and quantum
operations necessary.
Entanglement and enhanced security Quantum computing: The utilization of the concepts
Entanglement, a basic element of quantum com- of quantum physics is employed, wherein qubits
puting, refers to the situation in which the state of are utilized to represent several states concurrently.
one qubit is intrinsically connected to the state of Quantum computers possess the capability to tackle
another qubit, irrespective of the spatial separation specific issue classes at a significantly accelerated rate
between them. The aforementioned characteristic compared to traditional computers. In the field of
can be leveraged to develop encryption techniques cryptography, this phenomenon has a dual impact:
that are highly resistant to decryption, shown by an unparalleled challenge to existing encryption tech-
quantum key distribution (QKD) (Li et al., 2022). niques and the prospect of developing encryption
Quantum key distribution (QKD) protocols employ systems that are nearly impervious to decryption (Shi
entangled qubits as a means to securely disseminate and Li, 2022).
cryptographic keys. Any endeavor to intercept the
key exchange process results in the modification of Post-quantum cryptography (PQC) by NIST
the quantum state, hence exposing the existence of The primary objective of the National Institute of
an unauthorized party attempting to gain access. The Standards and Technology’s (NIST) post-quantum
incorporation of quantum entanglement into proto- cryptography (PQC) effort is to engage in the devel-
col design enables the attainment of an enhanced opment of cryptographic standards that possess the
level of security, a matter of utmost importance in resilience necessary to withstand the computational
the current period characterized by escalating cyber capabilities of quantum computers. The significance
threats and vulnerabilities. of this endeavor lies in the potential obsolescence of
540 Smart protocol design: Integrating quantum computing models
EncodedData = EncodeData(ClassicalData, QK) Data processing speed: This statistic demonstrates the
// Set up the quantum communication channel growth in computational capacity for data process-
QuantumChannel = SetUpQuantumChannel() ing. The use of quantum technology into the architec-
// Transmit encoded data ture yields substantial enhancements in performance,
TransmitData(EncodedData, QuantumChannel) resulting in accelerated computational processes and
// Distribute quantum key for decryption expedited data transmission.
QuantumKeyDistribution(QuantumChannel, Encryption strength: This statement reflects the level of
QK) strength and effectiveness exhibited by the encryption
// Receive and decode data employed. The protocol architecture has been improved
SecureDataTransmission = by integrating post-quantum techniques, resulting in
ReceiveData(Quantum Channel, QK) greater encryption capabilities that provide increased
Return SecureDataTransmission resistance against both classical and quantum attacks.
End Resource utilization: This statement highlights the
measure of effectiveness in utilizing computing
Results resources. Quantum computing models, renowned
for their high efficiency, effectively minimize resource
Table 68.2 compares key performance metrics of utilization on a global scale.
traditional protocol designs with those enhanced by
Latency: The reduced latency observed in the
quantum computing models and post-quantum cryp-
improved protocols serves as evidence for the efficacy
tography algorithms.
Table 68.4 Results analysis traditional protocol design and QC/PQC-enhanced protocol design.
of quantum models in both request processing and and quantum computing models. The integration of
data transmission. network protocols not only improves efficiency and
Error rate: A lower error rate in the quantum- processing capacities, but also provides a higher level
enhanced design points to the increased reliability and of security that is resistant to both conventional and
accuracy of the protocols. quantum computing threats. The incorporation of
PQC algorithms, as suggested by prominent organi-
Scalability: The quantum-enhanced design exhibits
zations like as the National Institute of Standards and
substantial improvements in its capacity to manage
Technology (NIST), strengthens this strategy by pro-
heightened workload and network expansion.
viding a structure that is resilient to future advance-
Security against quantum attacks: Conventional pro- ments and capable of accommodating changing
tocols typically lack the necessary capabilities to with- technological environments. Moreover, the investiga-
stand quantum attacks, whereas the improved design, tion of diverse quantum algorithms and their imple-
which integrates PQC, demonstrates robustness in the mentations in network security offers a glimpse into
face of these vulnerabilities. the prospective landscape of secure communications.
When these algorithms are included into network pro-
Efficiency gain pertains to the enhancement in pro- tocols, they provide a twofold benefit: they harness
cessing speed or resource utilization as compared to the computational capabilities of quantum computing
conventional algorithms (Table 68.3). to enhance efficiency, while also utilizing the resilience
Security enhancement measures refer to the imple- of PQC to ensure security. The aforementioned two-
mentation of strategies aimed at bolstering security, fold benefit holds significant importance in a contem-
particularly in the face of potential risks posed by porary context where the preservation and protection
quantum computing. of data integrity and security are of utmost signifi-
The concept of time complexity pertains to the cance. The incorporation of quantum computing
computer resources that are necessary for a certain models and PQA into protocol design is not merely a
algorithm. theoretical concept, but rather a pressing necessity for
Quantum robustness refers to the evaluation of the progression of network security. With the ongoing
an algorithm’s ability to withstand potential attacks advancement and increasing availability of quantum
originating from quantum computers. computing, it is imperative that the protocols we cur-
The limitations of each algorithm are identified to rently develop possess the necessary capabilities to
illustrate the challenges and drawbacks connected effectively address the cryptographic obstacles that
with them (Table 68.4). will arise in the future. This study thus presents a per-
suasive argument for researchers, technologists, and
Conclusion politicians to give utmost importance to the advance-
ment and adoption of quantum-resilient protocols, in
Our investigation culminates in the recognition that order to guarantee a future that is both secure and
the incorporation of quantum computing models into efficient in the realm of digital technology.
protocol design signifies a substantial advancement in
the domains of network security and efficiency. The
implementation of post-quantum algorithms (PQA) References
in this particular context signifies a significant shift Ji, Z., Zhang, H., and Wang, H. (2019). Quantum private
towards enhancing the security of digital commu- comparison protocols with a number of multi-particle
nications against existing and future cryptographic entangled states. IEEE Acc., 7, 44613–44621.
vulnerabilities, particularly in the age of quantum Anshu, Anurag, Shima Bab Hadiashar, Rahul Jain, Ashwin
Nayak, and Dave Touchette. (2023). One-shot quan-
computing. The study highlights the significant influ-
tum state redistribution and quantum Markov chains.
ence of quantum computing on conventional cryp- IEEE Transactions on Information Theory, 69, 5788–
tography protocols through its debates and analysis. 5804.
Quantum computing models provide exceptional Xu, RuiQing, Ri-Gui Zhou, and YaoChong Li. (2023). To-
computational speed and efficiency, hence facilitating wards the advantages of quantum trajectories on en-
the development of more resilient and secure network tanglement distribution in quantum networks. IEEE
protocols. Nevertheless, the emergence of quantum Transactions on Wireless Communications, 22, 5170–
computing also presents novel concerns, namely in 5184.
terms of the potential risks it poses to existing cryp- Li, J., Jia, Q., Xue, K., Wei, D. S. L., and Yu, N. (2022). A
tography protocols. The use of PQC is needed in order connection-oriented entanglement distribution design
to ensure security against the powerful capabilities of in quantum networks. IEEE Trans. Quan. Engg., 3,
1–13.
quantum computers. This work elucidates a synergistic
Chiti, F, Picchi, R., and Pierucci, L. (2022). Metropolitan
approach to protocol creation by incorporating PQC quantum-drone networking and computing: A soft-
Applied Data Science and Smart Systems 543
ware-defined perspective. IEEE Acc., 10, 126062– nism QDPoS. IEEE Trans. Inform. Forens. Sec., 17,
126073. 3264–3276.
Shi, R.-H. and Li, Y.-F. (2022). Quantum protocol for secure El-Latif, A., Ahmed, A., Abd-El-Atty, B., Shamim Hossain,
multiparty logical AND with application to multipar- M., Elmougy, S., and Ghoneim, A. (2018). Secure
ty private set intersection cardinality. IEEE Trans. Cir. quantum steganography protocol for fog cloud inter-
Sys. I Reg. Papers, 69(12), 5206–5218. net of things. IEEE Acc., 6, 10332–10340.
Shi, R.-H. (2020). Quantum multiparty privacy set intersec- Guo, C., Liang, F., Lin, J., Xu, Y., Sun, L., Liu, W., Liao, S.,
tion cardinality. IEEE Trans. Cir. Sys. II Exp. Briefs, and Peng, C. (2019). Control and readout software for
68(4), 1203–1207. superconducting quantum computing. IEEE Trans.
Cacciapuoti, A. S., Caleffi, M., Van Meter, R., and Hanzo, Nuc. Sci., 66(7), 1222–1227.
L. (2020). When entanglement meets classical com- Sehra, S. S., Singh, J., Rai, H. S., and Anand, S. S. (2020).
munications: Quantum teleportation for the quantum Extending processing toolbox for assessing the logi-
internet. IEEE Trans. Comm., 68(6), 3808–3833. cal consistency of OpenStreetMap data. Transac. GIS,
Doolittle, B., Thomas Bromley, R., Killoran, N., and Chit- 24(1), 44–71.
ambar, E. (2023). Variational quantum optimization Shi, R.-H. and Li, Y.-F. (2022). A feasible quantum sealed-
of nonlocality in noisy quantum networks. IEEE bid auction scheme without an auctioneer. IEEE
Trans. Quan. Engg., 4, 1–27. Trans. Quan. Engg., 3, 1–12.
DiAdamo, S., Nötzel, J., Zanger, B., and Beş e, M. M. (2021). Liu, Wen-Jie, and Zi-Xian Li. (2023). Secure and Efficient
Qunetsim: A software framework for quantum net- Two-Party Quantum Scalar Product Protocol With
works. IEEE Trans. Quan. Engg., 2, 1–12. Application to Privacy-Preserving Matrix Multipli-
Shi, R.-H. (2021). Anonymous quantum sealed-bid auction. cation. IEEE Transactions on Circuits and Systems I:
IEEE Trans. Cir. Sys. II Exp. Briefs, 69(2), 414–418. Regular Papers, 70, 4456–4469.
Shi, R.-H. and Li, Y.-F. (2022). Quantum secret permutating Van Meter, R., Ladd, T. D., Munro, W. J., and Nemoto, K.
protocol. IEEE Trans. Comp., 72(5), 1223–1235. (2008). System design for a long-line quantum repeat-
Yang, Z., Salman, T., Jain, R., and Di Pietro, R. (2022). De- er. IEEE/ACM Trans. Netw., 17(3), 1002–1013.
centralization using quantum blockchain: A theoreti- Yan, Z., Meyer-Scott, E., Bourgoin, J.-P., Higgins, B. L.,
cal analysis. IEEE Trans. Quan. Engg., 3, 1–16. Gigov, N., MacDonald, A., Hübel, H., and Jennewein,
Song, D. and Chen, D. (2020). Quantum key distribution T. (2013). Novel high-speed polarization source for
based on random grouping bell state measurement. decoy-state BB84 quantum key distribution over free
IEEE Comm. Lett., 24(7), 1496–1499. space and satellite links. J. Lightw. Technol., 31(9),
Li, Q., Wu, J., Quan, J., Shi, J., and Zhang, S. (2022). Ef- 1399–1408.
ficient quantum blockchain with a consensus mecha-
69 Efficient IIoT framework for mitigating Ethereum attacks
in industrial applications using supervised learning with
quantum classifiers
S. B. Goyal1,a, Anand Singh Rajawat2, Ritu Shandilya3 and Varun Malik4
Faculty of Information Technology, City University, Petaling Jaya, 46100, Malaysia
1
School of Computer Science & Engineering, Sandip University, Nashik, Maharashtra, India
2
Associate Professor of Computer and Data Science, Mount Mercy University, Cedar Rapids, Iowa, USA
3
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
4
Abstract
Industrial Internet of Things (IIoT) solutions have transformed industrial productivity and operations. The incorporation
of Ethereum blockchain technology into IIoT creates new weaknesses, exposing industrial systems to several cyberattacks.
An unique IIoT framework mitigates Ethereum-based attacks in industrial applications to solve these vulnerabilities. This
system uses supervised learning and quantum classifiers to detect and fix fraudulent Ethereum transaction patterns in real
time. Our methodology has lower false positive rates and higher detection accuracy than conventional methods, according
to first trials. This study shows that quantum computing and machine learning (ML) can improve the security of Ethereum-
enabled IIoT devices in industry.
Keywords: IIoT security, Ethereum attack mitigation, industrial applications, supervised learning, quantum classifiers, ef-
ficient framework
drsbgoyal@[Link]
a
Applied Data Science and Smart Systems 545
Table 69.1 Comparative analysis.
The focus of Huang et al. (2022) study is on the publications is provided. The tabular representation
tactics employed for representing data in the context below was constructed using the titles and citations
of process monitoring in IIoT. The proposed solution of the papers. To attain a comprehensive understand-
addresses the challenge of integrating the handling of ing of the subject matter, it is important to engage
both stationary and nonstationary data. The meth- in a thorough examination of all pertinent articles
odologies proposed in this study have the potential (Table 69.1).
to enhance manufacturing practices in the context of
IIoT. Proposed methodology
Sinha et al.’s (2022) article presents the concept
of “iThing,” which emphasizes the significance of Using supervised learning with quantum classi-
incorporating self-monitoring capabilities for battery fiers: An effective IIoT framework for protecting
health into the design of Internet of Things devices. against Ethereum attacks in high-stakes industrial
This endeavor contributes to our objective of guaran- environments.
teeing the longevity and reliability of IIoT devices, a
critical factor for their sustained sustainability. Data collection
The paper by Yang et al. (2019) introduces a novel The data collected includes information from vari-
framework called “IIoT-MEC” that leverages mobile ous sensors, network activity, and system logs that are
edge computing (MEC) to support 5G-enabled IIoT created by a diverse array of industrial applications
applications. Due to its capability of offering edge operating in real-time (Li et al., 2022).
processing with little latency, MEC is highly suitable Create a comprehensive repository of Ethereum
for the real-time processing and control of IIoT data. attack data encompassing instances of both success-
The paper by Liao et al. (2020) presents a compre- ful and unsuccessful attacks, various attack channels
hensive analysis of a demand response paradigm that employed, and discernible patterns.
incorporates computational intelligence in the context The study titled “Efficient IIoT framework for
of interaction networks between IIoT and MEC sys- mitigating Ethereum attacks in industrial applica-
tems. In order to optimize the efficacy of IIoT applica- tions using supervised (Yang and Shami, 2023) learn-
tions inside MEC environments, a key focus is placed ing with quantum classifiers” requires the creation
on efficiently allocating computing resources. of a dataset table. This table should encompass the
A tabular representation of the merits, flaws, oppor- properties, descriptions, and types of data repre-
tunities, and threats of the three aforementioned sented within the dataset. Table 69.2 derived from
546 Efficient IIoT framework for mitigating Ethereum attacks in industrial applications
Table 69.2 An illustrative dataset for a robust IIoT architecture.
Abstract
The incorporation of quantum computing within the framework of the Internet of Things (IoT) signifies a fundamental
transformation in the realm of data processing and security pertaining to interconnected devices. This study examines the
profound influence of quantum computing on IoT, with a specific emphasis on its capacity to fundamentally alter data
management practices and bolster security protocols. Quantum computing, renowned for its remarkable capacity to execute
intricate computations at unparalleled velocities, presents notable strides in computational capability and effectiveness. The
significance of this matter is particularly pronounced within IoT environment, because a multitude of devices collect and
exchange substantial volumes of data. This study explores the utilization of quantum algorithms for the efficient processing,
analysis, and security of data, specifically focusing on the challenges faced by classical computing in managing the vast scale
and intricate nature of IoT networks. Furthermore, this paper examines the distinctive features of quantum cryptography,
which offer resilient security measures against growing cyber threats. This is a crucial factor within the context of IoT envi-
ronment. This study also investigates the obstacles and possible remedies involved in the integration of quantum computing
with IoT devices, encompassing constraints related to hardware and scalability. This detailed analysis elucidates the potential
for quantum computing to bring about transformation in the realm of IoT networks, hence augmenting their capabilities
and security. Consequently, this development paves the way for the emergence of more advanced, efficient, and secures linked
devices across diverse sectors. To proposed the algorithm with the combination of adiabatic quantum computing (AQC) and
quantum key distribution (QKD).
Keywords: Quantum computing, Internet of Things (IoT), data processing, quantum cryptography, IoT security, scalability
drsbgoyal@[Link]
a
Applied Data Science and Smart Systems 553
streams is of utmost importance for applications that integration. Due to quantum computing technolo-
necessitate such functionality, including autonomous gies that challenge cryptography, it stresses the
vehicles and real-time environmental monitoring. necessity for strong security solutions. To secure
Moreover, quantum computing represents a funda- cyber systems, the research proposes hybrid tech-
mental transformation in the realm of data security, niques using classical and quantum-resistant
which is a matter of utmost importance in the con- algorithms.
text of IoT networks (Ahmad et al., 2021). The emer- A quantum tunneling physically unclonable func-
gence of quantum algorithms presents sophisticated tion (PUF) introduced as a new hardware security
cryptography methodologies, hence guaranteeing method in a research done by Chuang et al. (2021).
the establishment of safe communication protocols It suggests using quantum tunneling to create unique,
across various gadgets. The significance of this com- unclonable fingerprints for semiconductor chips to
ponent is growing in importance as IoT ecosystem prevent counterfeiting and tampering.
is frequently targeted by cyberattacks, mostly due The research by Al-Mohammed and Yaacoub
to its extensive range of uses and ease of access. In (2021) examines how quantum communication tech-
summary, the incorporation of quantum computing nologies can safeguard IoT devices in the 6G future.
within the domain of IoT holds the potential to effec- It addresses how quantum key distribution and
tively tackle the concurrent issues of data process- other quantum-based technologies can safeguard
ing and security. Through the utilization of quantum the growing IoT infrastructure against advanced
mechanics, there exists the potential to augment the cyberattacks.
efficacy, velocity, and safeguarding of data processing Shim (2021) in his survey examines post-quantum
in interconnected devices, thereby assuming a crucial public-key signature techniques for secure vehicle
function in the progression of IoT domain. In light of communications. It evaluates quantum-resistant
the current technological advancements, it is crucial cryptography methods for intelligent transportation
to thoroughly investigate and exploit the capabilities systems.
of quantum computing in order to effectively achieve Each of these works advances quantum technolo-
the potential of IoT era. gies and their applications in communication, secu-
“Explore integration of Adiabatic Quantum rity, and IoT, shedding light on the difficulties and
Computing and Quantum Key Distribution in IoT for solutions of a quickly changing quantum-influenced
enhanced data security.” technological landscape.
“Develop algorithms combining AQC and QKD to
revolutionize IoT device communication and encryp- Purposed methodology
tion protocols.”
“Investigate synergies between AQC and QKD The incorporation of quantum computing (QC)
to significantly improve IoT network security and (Sandilya and Sharma, 2021) inside the framework
efficiency.” of IoT represents a significant advancement in the
This paper is organized as to represent the related realms of data processing and security. The IoT is
work, proposed methodology, results analysis, and distinguished by its extensive network of intercon-
finally conclusion and future work. nected devices, which results in the generation of
substantial amounts of data. Consequently, the pro-
cessing of this data requires advanced computational
Related work
skills, as well as the implementation of solid security
The related works you mentioned cover a range of measures. Quantum computing presents a promis-
topics in the fields of quantum communication, cyber- ing avenue for addressing these difficulties, given its
security in the quantum era, and applications of quan- remarkable computational capabilities and promise
tum technologies in various domains. The following is for unmatched levels of security.
a summary of each work:
The study by Sandilya and Sharma (2021) explores Quantum computing: A paradigm shift
the quantum internet and its potential to transform Quantum computing utilizes the fundamental prin-
global communications. We examine the technologi- ciples of quantum mechanics, employing qubits that
cal advances and obstacles of building a quantum have the ability to exist in superposition, allowing
internet infrastructure. They focus on how entangle- for simultaneous occupation of multiple states. This
ment and quantum key distribution might create enables quantum computers to execute intricate cal-
unprecedented security and efficiency in quantum culations at velocities that cannot be achieved by con-
communication. ventional computers. In the realm of IoT, quantum
The paper by Yavuz et al. (2022) examines post- computing possesses the capability to efficiently han-
quantum distributed cyber-infrastructures and AI dle substantial datasets, hence enabling the possibility
554 Quantum computing in the era of IoT
energy, which corresponds to the most optimal solu- to the improvement of security algorithms. For
tion. This functionality is especially advantageous example, the rapid problem-solving capabilities of
for activities such as pattern identification, anomaly this technology can be leveraged to enhance encryp-
detection, and prognostic maintenance in IoT devices tion algorithms, hence enhancing their resistance to
(Malina et al., 2021). both classical and quantum attacks. This holds special
significance within the framework of devising quan-
Proposed Algorithm 2 tum-resistant encryption techniques for IoT devices
Algorithm: AQC_IoT_Data_Processing (Aminanto et al., 2017; Xin et al., 2020).
Inputs:
IoT_Data: Multidimensional data from IoT Proposed algorithm 3
devices Algorithm: Enhance_IoT_Security_with_AQC
Optimization_Problem: The specific optimiza- Input: IoT_Device_Data, Classical_Encryption_
tion problem to be solved Parameters
Output: Output: Quantum_Safe_Encrypted_Data
Optimal_Solution: The best solution found for Procedure Enhance_IoT_Security_with_AQC:
the given problem // Step 1: Initialize IoT device data and encryp-
Begin tion parameters
// Initialize the quantum system device_data <- IoT_Device_Data
Quantum_System classical_parameters <- Classical_Encryption_
<- Initialize_Quantum_System() Parameters
// Map the optimization problem onto the quan- // Step 2: Define AQC optimization problem for
tum system’s energy landscape encryption
Energy_Landscape <- Map_Problem_To_Energy_ Define AQC_Optimization_Problem:
Landscape(IoT_Data, Optimization_Problem) Objective: Minimize the potential of the system
// Set the initial and final Hamiltonians Constraints: Adhere to quantum mechanics
Initial_Hamiltonian <- Define_Initial_ principles
Hamiltonian (Energy_Landscape) // Step 3: Encode the encryption problem into
Final_Hamiltonian <- Define_Final_Hamiltonian AQC
(Energy_Landscape) aqc_problem <- Encode_Encryption_Problem
// Set the total evolution time (device_data, classical_parameters)
Total_Time <- Define_Total_Evolution_Time() // Step 4: Solve the optimization problem using
// Apply adiabatic evolution AQC
For t from 0 to Total_Time do optimized_solution <- Solve_AQC_Optimization
Current_Hamiltonian <- Adiabatic_Evolution _Problem(aqc_problem)
(Initial_Hamiltonian, Final_Hamiltonian, t, // Step 5: Extract quantum-safe encryption
Total_Time) parameters
quantum_safe_parameters <- Extract_Parameters
Update_Quantum_System_State(Quantum_ (optimized_solution)
System, Current_Hamiltonian) // Step 6: Encrypt IoT device data using quantum-
End safe parameters
// Measure the quantum system to obtain the Quantum_Safe_Encrypted_Data <-
solution Encrypt(device_data, quantum_safe_parameters)
Optimal_Solution <- Measure_Quantum_System // Step 7: Return the encrypted data
(Quantum_System) return Quantum_Safe_Encrypted_Data
End Procedure
Return Optimal_Solution
End AQC for energy-efficient IoT operations
Internet of Things (IoT) devices frequently functions
Enhancing IoT security with AQC within limitations pertaining to power consumption
The issue of security in IoT (Rahman et al., 2017) is and processing capabilities. Adiabatic quantum com-
of utmost importance, particularly due to the wide- puting (AQC) exhibits a more energy-efficient nature
spread adoption of devices and the high level of sensi- in comparison to classical computing techniques and
tivity associated with the data being transmitted and alternative quantum computing approaches, owing to
stored. Automated query construction (AQC) (Tulli et its slow and regulated generation of quantum states.
al., 2019) has the potential to significantly contribute The inclusion of this functionality is crucial in order
556 Quantum computing in the era of IoT
Table 70.1 Comparative analysis between traditional IoT system and quantum-enhanced IoT system.
DS1 Processing speed (Ops/sec) 1000 Ops/sec 5000 Ops/sec 400% DS1
DS2 Data throughput (GB/hr) 50 GB/hr 200 GB/hr 300% DS2
DS3 Encryption strength (bit) 128-bit 256-bit (quantum Enhanced DS3
resistant)
DS4 Error rate (%) 2% 0.5% Reduced DS4
DS5 Energy efficiency (Joules/ 0.01 Joules/Op 0.005 Joules/Op 50% Saving DS5
Op)
DS6 Latency (ms) 10 ms 2 ms 80% Reduced DS6
DS7 Scalability (Max devices) 10,000 devices 50,000 devices 400% DS7
DS8 Network resilience (Score) 3/5 5/5 Improved DS8
Applied Data Science and Smart Systems 557
Table 70.2 Results analysis between traditional IoT system and quantum-enhanced IoT system.
Data processing Limitations of classical Speedier due to quantum Quantum computing processes
speed computation superposition and parallelism data almost instantly, surpassing
regular methods
Security Quantum assaults may Quantum assaults may Quantum cryptography
measures compromise encryption compromise encryption provides impenetrable security,
even for quantum computers
Data handling Limited by classical computing Exponentially increased due to IoT devices capture massive
capacity qubits’ enhanced information volumes of data, hence quantum
capacity computing is necessary
Energy efficiency Classical computing uses more Quantum processing efficiency Although speculative, quantum
energy may reduce energy use computing could improve IoT
data processing energy efficiency
Scalability Limited computational and Compact quantum devices Quantum computing may let
energy resources limit scalability and efficient data processing IoT networks scale without
improve scalability resource limits
Response time to It uses traditional algorithms Quantum-based decryption Quantum-enhanced IoT systems
security threats and networks, making it slower and threat detection techniques detect and respond to security
provide fast reaction threats faster
capacity, distinguished by its capability to execute quantum computing. 2021 28th IEEE Int. Conf. Elec.
intricate computations at velocities beyond the reach Cir. Sys. (ICECS), 1–5.
of traditional computing. The capacity to perform this Sandilya, N. and Sharma, A. K. (2021). Quantum In-
function is of utmost importance within the IoT ecosys- ternet: An approach towards global communica-
tion. 2021 9th Int. Conf. Reliab. Infocom Technol.
tem, since it is characterized by the generation of huge
Optim. (Trends and Future Directions)(ICRITO),
volumes of data from numerous devices. Quantum
1–5.
computing possesses the capability to perform more Yavuz, A. A., Nouma, S. E., Hoang, T., Earl, D., and Pack-
efficient data analysis, hence facilitating real-time ard, S. (2022). Distributed cyber-infrastructures and
processing and decision-making. This attribute is artificial intelligence in hybrid post-quantum era.
vital for a wide array of applications, spanning from 2022 IEEE 4th Int. Conf. Trust Priv. Sec. Intel. Sys.
smart cities to personalized healthcare. Furthermore, Appl. (TPS-ISA), 29–38.
the ramifications for security in IoT resulting from Chuang, K. K.-H., Chen, H.-M., Wu, M.-Y., Ching-Sung
the advent of quantum computing are significant. The Yang, E., and Ching-Hsiang Hsu, C. (2021). Quan-
existing security mechanisms for IoT heavily rely on tum tunneling PUF: A chip fingerprint for hardware
traditional cryptographic approaches, which are sus- security. 2021 Int. Sympos. VLSI Technol. Sys. Appl.
(VLSI-TSA), 1–2.
ceptible to quantum assaults. Nevertheless, quantum
Al-Mohammed, H. A. and Yaacoub, E. (2021). On the use
computing presents a potential remedy in the form of
of quantum communications for securing IoT devices
quantum encryption techniques such as QKD, which in the 6G era. 2021 IEEE Int. Conf. Comm. Work-
holds the promise of offering security that is theoreti- shops (ICC Workshops), 1–6.
cally impervious to decryption. Ensuring the protec- Shim, K.-Ah. (2021). A survey on post-quantum public-key
tion of sensitive data transmitted across IoT networks signature schemes for secure vehicular communica-
is of utmost importance, as it safeguards privacy and tions. IEEE Trans. Intel. Trans. Sys., 23(9), 14025–
maintains data integrity within a context where secu- 14042.
rity breaches can yield significant and wide-ranging Alkhulaifi, A. and El-Alfy, E.-S. M. (2020). Exploring
ramifications. However, there are still obstacles that lattice-based post-quantum signature for JWT au-
need to be overcome in order to fully harness the thentication: review and case study. 2020 IEEE
91st Vehicul. Technol. Conf. (VTC2020-Spring),
capabilities of quantum computing in the context
1–5.
of IoT. The challenges encompassed in this domain
Althobaiti, O. S. and Dohler, M. (2020). Cybersecurity chal-
encompass technological barriers in the advancement lenges associated with the Internet of Things in a post-
of quantum systems that can be scaled effectively, the quantum world. IEEE Acc., 8, 157356–157381.
need to address energy efficiency concerns, and the Nikiema, P. R., Palumbo, A., Aasma, A., Cassano, L., Kri-
requirement for smooth interaction with established tikakou, A., Kulmala, A., Lukkarila, J., Ottavi, M.,
IoT infrastructures. Furthermore, as the technological Psiakis, R., and Traiola, M. (2023). Towards depend-
advancements progress, there will be a rise in neces- able RISC-V cores for edge computing devices. 2023
sity for a proficient labor force proficient in harness- IEEE 29th Int. Symp. On-Line Test. Robust Sys. Des.
ing its potential, as well as for regulatory frameworks (IOLTS), 1–7.
to effectively govern its consequences. In summary, Lee, W.-K., Jang, K., Song, G., Kim, H., Hwang, S. O.,
and Seo, H. (2022). Efficient implementation of
quantum computing is positioned to assume a pivotal
lightweight hash functions on gpu and quantum
function in the next era of IoT, presenting remedies for
computers for iot applications. IEEE Acc., 10,
a range of urgent issues pertaining to data processing 59661–59674.
and security. With the ongoing progress in research Singh, S., Singh, J., Goyal, S. B., Sehra, S. S., Ali, F., Alkha-
and development within this domain, it is anticipated faji, M. A., and Singh, R. (2023). A novel framework
that there will be a significant and profound influ- to avoid traffic congestion and air pollution for sus-
ence on the functioning and interaction of networked tainable development of smart cities. Sustain. Ener.
devices. This progress will facilitate the emergence of Technol. Assess., 56, 103125.
a more streamlined, secure, and interconnected global Malina, L., Dzurenda, P., Ricci, S., Hajny, J., Srivastava, G.,
landscape. Matulevič ius, R., Affia, A.-A. O., Laurent, M., Sultan,
N. H., and Tang, Q. (2021). Post-quantum era privacy
protection for intelligent infrastructures. IEEE Acc.,
References 9, 36038–36077.
Jang, G., Kim, D., Lee, I.-H., and Jung, H. (2023). Coop- Rahman, S. S., Heartfield, R., Oliff, W., Loukas, G., and
erative beamforming with artificial noise injection for Filippoupolitis, A. (2017). Assessing the cyber-trust-
physical-layer security. IEEE Acc., 11, 22553–22573. worthiness of human-as-a-sensor reports from mobile
Ahmad, S. F., Ferjani, M. Y., and Kasliwal, K. (2021). En- devices. 2017 IEEE 15th Int. Conf. Softw. Engg. Res.
hancing security in the industrial IoT sector using Manag. Appl. (SERA), 387–394.
Applied Data Science and Smart Systems 559
Tulli, D., Abellan, C., and Amaya, W. (2019). Engineer- Aminanto, M. E., Choi, R., Tanuwidjaja, H. C., Yoo, P. D.,
ing high-speed quantum random number generators. and Kim, K. (2017). Deep abstraction and weighted
2019 21st Int. Conf. Trans. Optic. Netw. (ICTON), feature selection for Wi-Fi impersonation detection.
1–1. IEEE Trans. Inform. Foren. Sec., 13(3), 621–636.
Xin, G., Han, J., Yin, T., Zhou, Y., Yang, J., Cheng, X., and Kumar, A., Kumar, S., Kumar, V., Kumari, A., Saini, A., and
Zeng, X. (2020). VPQC: A domain-specific vector Gupta, S. (2023). Edge computing based IDS detect-
processor for post-quantum cryptography based on ing threats using machine learning and PyCaret. 2023
RISC-V architecture. IEEE Trans. Cir. Sys. I Reg. Pa- Int. Conf. Comput. Intel. Sustain. Engg. Sol. (CISES),
pers, 67(8), 2672–2684. 668–673.
71 A federated learning approach to classify depression using
audio dataset
Chetna Gupta and Vikas Khullara
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
Vocal emotions are basic to expressing and understanding thoughts and the low toned vocal emotion expression is a major
deficit in individuals with depression. The primary objective of this paper is to propose a federated learning (FL)-based clas-
sification model to identify depression in individuals through audio. The current scenario of artificial intelligence (AI) focuses
on collaborative training of deep learning (DL) models without losing data privacy. So, in the methodology of this paper,
a collaborative and privacy preserved approach has been developed using FL for training deep learning models. The long-
short short-term memory (LSTM) and bidirectional- long-short term memory (B-LSTM)-based deep learning models will be
trained on a collected dataset in the federated learning ecosystem. As a result, the implemented models will be comparatively
analyzed on base DL structure as well as the FL ecosystem. The purpose of the investigation is to compare the impact of FL
architecture implementations on benchmark models. In conclusion, the most effective examined strategy will be considered
for future research objectives.
Keywords: Depression, artificial intelligence, federated learning, long-short term memory, bidirectional – long-short term
memory, audio
[Link]@[Link]
a
Applied Data Science and Smart Systems 561
improve various health research. Pranto and Al Asad the audio waves are then split across four clients for
(2021) introduced the application of FL in medical IID settings. A vocal descriptor that is frequently
studies by comparing centralized and FL approaches used to identify depression is Mel Frequency Cepstral
across multiple mental diseases. Khullar and Singh Coefficients (MFCC). Consequently, 162 features have
(2022) introduced FL trained algorithms for identify- been retrieved from the audio dataset using MFCC.
ing disaster areas in the internet of unmanned aerial To see the outcomes in many scenarios with the
vehicles (UAVs) that enhance data sharing. Fan et al. highest performance, DL algorithms are implemented
(2021), Uyulan et al. (2021) and Yasin et al. (2021) in the audio file after pre-processing information. In
developed diagnostic systems to detect major depres- order to classify depressed and non-depressed par-
sive disorder using various DL techniques. ticipants, DL algorithms such as LSTM, CNN, and
Bi-LSTM were implemented. Moreover, these algo-
Methodology rithms are used to develop a base model and find the
best algorithm to work with FL.
In this study, an openly available audio dataset is Table 71.1 and Figure 71.2 shows the results of
obtained which are pre-processed and analyzed using base DL model algorithms in which using CNN,
deep learning algorithms as a base, CNN, long-short LSTM, and Bi-LSTM validation results reached 85%,
short-term memory (LSTM) and bidirectional- long- 89%, and 91%, respectively. So, the Bi-LSTM algo-
short term memory (B-LSTM). Following that, the rithm outperformed other algorithms with 91% high-
privacy-preserved FL algorithm is applied to IID est validation accuracy.
users to train a centralized model on the client site The FL method and Bi-LSTM method are combined
and develop an aggregated model on the server site. to analyze the IID data for the purpose of diagnosing
The audio files contain 52 individuals’ data from depressed patients while maintaining their privacy.
Chinese people (23 depressed and 29 healthy par- For the 4 clients, the training and validation sets are
ticipants) obtained from Lanzhou University’s Second divided as IID data. Further, all client data is tested
Affiliated Hospital (Cai et al., 2020). Each individual at the server site to aggregate all the client models.
has 29 clips in this dataset, which are categorized as Finally, the server developed model is updated at all
positive, neutral, or negative emotional stimuli. The clients’ sites.
voice data is collected using high-quality equipment.
Participants in both the healthy and depressed groups
range in age from 18 to 55 years. Table 71.1 Training and validation results of DL models
FL clients simply share the outcomes of calculated for depression detection
weights to form an aggregated analysis model. To Parameters Bi LSTM CNN LSTM
protect data privacy, no data is transferred between
nodes. FL supports N clients (C1, C2, ... CN) with Accuracy 99.08333 99.66667 99.33333
datasets (D1, D2, ... DN). So, FL trains each client’s Validation 91 85 89
data independently to establish a decentralized deep accuracy
learning model (Figure 71.1). Precision 99.08333 99.66667 99.33333
Validation 91 85 89
(1) precision
Recall 99.08333 99.66667 99.33333
Results and discussion Validation 91 85 89
recall
The DL architecture is applied as a base in this research Loss 0.032376 0.012926 0.015737
to evaluate audio recordings from depressed and nor-
Validation loss 0.349078 0.439554 0.347897
mal individuals. Using the privacy protected FL system
Figure 71.1 Federated learning model for depression detection using audio data
562 A federated learning approach to classify depression using audio dataset
Figure 71.2 Training and validation accuracy and loss results for depression detection using Bi-LSTM, CNN and LSTM
algorithms
According to Table 71.2 and Figure 71.3–71.5, Table 71.2 Training and validation results of FL models
privacy protected FL model achieved 94.25% for depression detection using IID data
accuracy for training IID data over the client site.
Parameters Client IID Client IID Server IID
Further, validation accuracy for IID data at the cli- training validation validation
ent site is slightly lower at 85.66% and validation
accuracy for IID data is 86.66% at the server site. Accuracy 94.25 85.66667 86.66667
Although the FL model has a little less accuracy Precision 94.25 85.66667 86.66667
than the DL model for privacy protected systems Recall 94.25 85.66667 86.66667
it’s really needed. Because hospital’s data is pri-
Loss 0.160396 0.356732 0.32788
vate and many patients don’t want to disclose their
Figure 71.3 Client training results for depression detection using IID data
Applied Data Science and Smart Systems 563
Figure 71.4 Client validation results for depression detection using IID data
Figure 71.5 Server validation results for depression detection using IID data
information so, FL is the best way to train a secure After that, an automated privacy-protected FL-based
diagnosis system. depression detection framework is proposed using an
audio dataset. Therefore, the suggested FL framework
Conclusion obtained accuracy for IID data of 94.25% during
training 85.66% during validation on the client site,
In this article, we introduce a system for identifying and 86.66% validation accuracy on the server site.
and categorizing depressed and healthy individu- Thus, in the future, this method could be applied to
als from audio data. The creation of an automated diverse datasets to check its reliability and robustness.
diagnostic system will aid clinicians in providing a
fast diagnosis of depression. So, firstly we developed
References
a base DL model using CNN, LSTM, and Bi-LSTM
algorithms in which Bi-LSTM outperformed other Adarsh, V., P. Arun Kumar, V. Lavanya, and G. R. Gangad-
algorithms with the highest 91% validation accuracy. haran. (2023). Fair and explainable depression detec-
564 A federated learning approach to classify depression using audio dataset
tion in social media. Information Processing & Man- Mousavian, M., Chen, J., Traylor, Z., and Greening, S.
agement, 60(1): 103168 [Link] (2021). Depression detection from SMRI and Rs-
ipm.2022.103168. FMRI images using machine learning. J. Intel. Inform.
Cai, Hanshu, Yiwen Gao, Shuting Sun, Na Li, Fuze Tian, Sys., 57(2), 395–418. [Link]
Han Xiao, Jianxiu Li et al. (2020). Modma dataset: a 021-00653-w.
multi-modal open dataset for mental-disorder analy- Orabi, Ahmed Husseini, Prasadith Buddhitha, Mahmoud
sis. arXiv preprint arXiv:2002.09283 [Link] Husseini Orabi, and Diana Inkpen. (2018). Deep
org/10.48550/arXiv.2002.09283. learning for depression detection of twitter users. In
Cai, Hanshu, Jiashuo Han, Yunfei Chen, Xiaocong Sha, Proceedings of the fifth workshop on computational
Ziyang Wang, Bin Hu, Jing Yang et al. (2018). A per- linguistics and clinical psychology: from keyboard to
vasive approach to EEG-based depression detection. clinic, 88–97.
Complexity 2018: 1–13. Pranto, Md A. M., and Al Asad, N. (2021). A comprehen-
Cui, Y., Li, Z., Liu, L., Zhang, J., and Liu, J. (2022). Pri- sive model to monitor mental health based on feder-
vacy-preserving speech-based depression diagnosis ated learning and deep learning. Proc. 2021 IEEE
via federated learning. Proc. Ann. Int. Conf. IEEE Int. Conf. Sig. Proc. Inform. Comm. Sys. SPICSCON
Engg. Med. Biol. Soc. EMBS, 1371–1374. Institute of 2021, 18–21. Institute of Electrical and Electron-
Electrical and Electronics Engineers Inc. [Link] ics Engineers Inc. [Link]
org/10.1109/EMBC48229.2022.9871861. SCON54707.2021.9885430.
Elbeltagi, A., Aslam, M. R., Malik, A., Mehdinejadiani, Sadilek, Adam, Luyang Liu, Dung Nguyen, Methun Kam-
B., Srivastava, A., Bhatia, A. S., and Deng, J. (2020). ruzzaman, Stylianos Serghiou, Benjamin Rader, Alex
The impact of climate changes on the water footprint Ingerman et al. (2021). Privacy-first health research
of wheat and maize production in the Nile delta, with federated learning. NPJ digital medicine. 4(1):
Egypt. Sci. Total Environ., 743, 140770. [Link] 132 pp. 1–8.
org/10.1016/[Link].2020.140770. Statista, R. D. (2021). Number of suicides India 1972–2019.
Fan, Z., Su, J., Gao, K., Peng, L., Qin, J., Shen, H., Hu, D., [Link]
and Zeng, L.-L. (2021). Federated learning on struc- of-suicides-india/.
tural brain MRI scans for the diagnostic classification Suruliraj, Banuchitra, and Rita Orji. (2022). Federated
of major depression. Biol. Psych., 89(9), S183. www. Learning Framework for Mobile Sensing Apps in
[Link]/journal. Mental Health. In 2022 IEEE 10th International Con-
Gupta, Chetna, and Vikas Khullar. (2022). Contemporary ference on Serious Games and Applications for Health
Intelligent Technologies for Electroencephalogram- (SeGAH), 1–7. IEEE.
Based Brain Computer Interface. In 2022 10th Inter- Ramesh, T. R., Lilhore, U. K., Poongodi, M., Simaiya, S.,
national Conference on Reliability, Infocom Technolo- Kaur, A., and Hamdi, M. (2022). Predictive analysis
gies and Optimization (Trends and Future Directions) of heart diseases with machine learning approach-
(ICRITO), 1–6. IEEE. es. Malaysian J. Comp. Sci., 132–148. [Link]
Khullar, Vikas, Raj Gaurang Tiwari, Ambuj Kumar Agar- org/10.22452/mjcs.sp2022no1.10.
wal, and Soumi Dutta. (2022). Physiological signals Uyulan, C., Ergüzel, T. T., Unubol, H., Cebi, M., Sayar, G.
based anxiety detection using ensemble machine H., Asad, M. N., and Tarhan, N. (2021). Major de-
learning. In Cyber Intelligence and Information Re- pressive disorder classification based on different
trieval: Proceedings of CIIR 2021, 597–608. Springer convolutional neural network models: Deep learning
Singapore. approach. Clin. EEG Neurosci., 52(1), 38–51. https://
Khullar, Vikas, and Harjit Pal Singh. (2022). Privacy pro- [Link]/10.1177/1550059420916634.
tected internet of unmanned aerial vehicles for disas- Yasin, Sana, Syed Asad Hussain, Sinem Aslan, Imran
trous site identification. Concurrency and computa- Raza, Muhammad Muzammel, and Alice Othmani.
tion: practice and experience. 34(19): e7040 https:// (2021). EEG based Major Depressive disorder and
[Link]/10.1002/cpe.7040. Bipolar disorder detection using Neural Networks:
Brar, P. S., Shah, B., Singh, J., Ali, F., and Kwak, D. (2022). A review. Computer Methods and Programs in Bio-
Using modified technology acceptance model to evalu- medicine, 202: 106007 [Link]
ate the adoption of a proposed IoT-based indoor disas- cmpb.2021.106007.
ter management software tool by rescue workers. Sen- Ye, J., Yu, Y., Wang, Q., Li, W., Liang, H., Zheng, Y., and
sors, 22(5), 1866, [Link] Fu, G. (2021). Multi-modal depression detection
LeMoult, J. and Gotlib, I. H. (2019). Depression: A cogni- based on emotional audio and evaluation text. J. Af-
tive perspective. Clin. Psychol. Rev., 69, 51–66. https:// fec. Disord., 295, 904–913. [Link]
[Link]/10.1016/[Link].2018.06.008. jad.2021.08.090.
72 Securing IOT CCTV: Advanced video encryption
algorithm for enhanced data protection
Kawalpreet Kaur1,2,a, Amanpreet Kaur3, Vidhyotma Gandhi4 and
Bhupendra Singh5
1,3
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
2
Goswami Ganesh Dutta Sanatan Dharma College, Sector 32, Chandigarh, India
4
Gyancity Research Labs, Gurugram, Haryana, India
5
Defence Research and Development Organization, Bangalore, Karnataka, India
Abstract
Ensuring the security of multimedia content and its applications has evolved into a pivotal responsibility within IoT com-
munication technology. Cryptography stands out as a fundamental technique that offers security and confidentiality, thereby
thwarting unauthorized data access. Choosing an appropriate design for an encryption system becomes imperative to guar-
antee comprehensive security and the preservation of data privacy. This study focuses on IoT-centric CCTV (CCTV) sys-
tems, illuminating potential security weaknesses during data transmission and storage. Through experimental assessments,
we evaluate these algorithms’ performance and security characteristics, considering factors such as encryption time, and
computational efficiency. By providing insights into the strengths and weaknesses of various video encryption methods, this
research aids in the implementation of robust security measures for IoT-driven surveillance applications. In this paper, an
optimized video encryption algorithm (OVEA) is proposed. In order to validate the effectiveness of the proposed algorithm,
a comprehensive set of experiments is conducted. A comparative analysis is carried out against existing encryption methods
to do the relative assessment The proposed OVEA algorithm demonstrates reduced encryption time of 0.00356s compared
to various existing public key algorithms when applied to video data.
a
kaur.kawalpreet17@[Link]
566 Securing IOT CCTV: Advanced video encryption algorithm for enhanced data protection
this study aims to provide insights into the critical balanced according to the level of risks involved. That
role of encryption algorithms in the security of IoT simply means it is difficult to recognize the face if the
CCTV cameras, shedding light on their effectiveness degree of masking is higher, thus enabling the stron-
in safeguarding sensitive video data and proposing an ger protection of data. According to Kim et al. (2020)
optimized video encryption algorithm (OVEA) that spoofing, sniffing, and inside attacks can be avoided
will optimize the encryption time and provide an opti- with this.
mal solution for the security of IoT CCTV devices. In an IoT surveillance system, a lot of video data is
generated that contains a large amount of insignifi-
Related work cant data. Therefore, to secure the useful data, Priya
et al. (2021) use a video summarization technique
In this segment, various relevant studies concern- that is used to extract meaningful frames from large
ing video encryption are explored and elucidated. video data to detect abnormal events. To detect the
These works delve into the examination and explana- abnormal image, feature extraction is done using the
tion of the security performance of video encryption Blob analysis method. After feature extraction, clas-
approaches in various research studies. Block cipher sification is performed to compare databases with the
encryption algorithms such as AES (Heron, 2009) detected objects using the K-NN algorithm. In the last,
which is appropriate for encrypting text data cannot encryption is done using the AES algorithm and then
be applied to the encryption of video streams due to an encrypted image is sent to the user using Gmail.
the low power processor. The problems of block cipher In IoT based home monitoring systems, surveillance
based encryption are solved using a permutation-based systems are used but to ensure privacy, lightweight
encryption algorithm that is proposed by Liu and security algorithms are used. To ensure security, data
Koenig (2005) and Gera et al. (2021). The video frame encryption is done using the keccak-chaotic sequence
here is encrypted by changing the order of one specific by Ravikumar and Kavita (2020).
part with another one in the frame. This algorithm is Hameed Obaida et al. (n.d.) reviewed video encryp-
suitable for video data encryption as it generates lower tion techniques for content protection, highlighting
processing overhead than AES. But because the same their vulnerabilities to cryptanalysis attacks. Fully
permutation list is being used for every frame, it is encryption techniques offer high security but are com-
unsafe for plain text attacks and as encryption is done putationally expensive and not suitable for real-time
after compression, hackers can still recover parts of use. Selective encryption algorithms are faster but
the original frame. Therefore, Sultana and Shubhangi provide lower video security.
(2017) proposed an encryption algorithm that encrypts Al-Husainy and Al-Shargabi (n.d.) discussed that a
video streams before compression and that is based on lightweight encryption model is required to combat
the faro shuffle algorithm. But still, plaintext attack the limited resources such as process and memory of
problems exist in this algorithm, as it uses the same IoT devices. Therefore, this model also ensures high-
permutation list of every frame, and since the complex- level security by using a large size key that is difficult
ity of the faro shuffle algorithm is low, it is unsafe from to crack, and that key is further changed after a cer-
brute force attacks. Therefore, Yun and Kim (2020) tain period (Table 72.1).
proposed an algorithm to avoid plain text attacks by
updating the permutation list for each frame. However,
Methodology
this algorithm increases the encryption time.
CCTV becomes the essential requirement to iden- Prominent issues associated with CCTV camera secu-
tify an individual based on their facial characteris- rity would be studied, to propose a cryptography
tics. But despite using deep learning, it is difficult to algorithm for the optimization of the time of CCTV
recognize faces correctly due to low resolution, acute footage of IoT devices. The performance of the pro-
weather, or various facial expressions. Therefore, Kim posed algorithm would be examined as per estab-
et al. (2020) proposed an access control technique lished parameters. The algorithm for the proposed
that is based on video surveillance. This technique is OVEA is given below:
used to incorporate CCTV machine learning in facial
recognition systems with radio frequency identifica- Capture the original video from CCTV camera.
tion (RFID) features that enable multichannel authen- Compress the original video using a video codec (e.g.,
tication on the mobile of the user. H.264, H.265, VP9) to reduce its file size.
If somehow these RFID authentication tags are Decompose the compressed video into individual
breached or in case of poor video quality, this dual frames for further processing.
channel authentication approach will be able to pro- Generate encryption keys for securing the video data.
tect the privacy of the user. Differential face image This involves symmetric encryption keys for the
masking is implemented in this approach that can be video frames.
Applied Data Science and Smart Systems 567
Table 72.1 Key findings in the literature
Heron, 2009 Block cipher encryption Text encryption Not applied in videos
Liu and Koenig, 2005 Permutation based encryption Suitable for video Vulnerable to plain text
encryption attacks
Sultana and Shubhangi, Faro shuffle algorithm Suitable for video Vulnerable to plain text and
2017 encryption brute force attacks
Yun and Kim, 2020 Permutation based encryption Safe from plain text attacks Encryption time increased
Kim et al., 2020 Access control technique and Prevent spoofing, sniffing Difficult to recognize the
differential face image masking and inside attacks face if the degree of masking
is higher
Priya et al., 2021 Video summarization technique Extract meaningful frames Not applied in real time
from large video data to
detect abnormal events
Hameed Obaida, n.d. Keccak-chaotic sequence Lightweight security Not used for different
algorithm formats of data
Al-Husainy and Lightweight encryption model Overcome the problem of -
Al-Shargabi, n.d. limited resources of IoT
Table 72.2 A comparison results of encryption speed time (in frame/seconds) for different videos without compression
CCTV Video size Video Frame rate/ Frame Encryption time per frame in Decryption time per frame in
video in KB length sec count sec of proposed algorithm sec of proposed algorithm
Encrypt each video frame using the generated encryp- without compressing CCTV footage are being com-
tion keys. This step ensures the data remains puted and results are shown in Table 72.2. Here,
confidential and secure. encryption is applied directly to the CCTV footage
Combine the encrypted frames to form an encrypted and encryption and decryption time is being calcu-
video sequence. lated. The proposed OVEA algorithm is giving dif-
Depending on the application, you can either display ferent encryption and decryption times according to
the encrypted video or transmit it securely to the video length, size of the video and frame count,
cloud storage via the internet for remote access. but giving optimal results as shown in Table 72.3 as
Retrieve the encrypted video from storage or recep- the video is first being compressed and then encrypted
tion over the internet. and video is uncompressed and decryption process
Decrypt each frame of the encrypted video using the is carried out. Figures 72.1 and 72.2 illustrate that
same encryption keys used for encryption. employing the proposed OVEA for encryption yields
optimal outcomes, thereby strengthening security
Results and discussion measures.
After conducting a comprehensive evaluation of
For the encryption process to be carried out, five dif- various encryption algorithms, it is evident that our
ferent CCTV videos from different cameras with vari- algorithm has emerged as the most efficient in terms
able size and frame count are being considered. The of encryption time as displayed in Table 72.4. The
encryption and decryption time of different videos encryption process, when executed using OVEA
568 Securing IOT CCTV: Advanced video encryption algorithm for enhanced data protection
Table 72.3 A comparison results of encryption speed time (in frame/seconds) for different videos with compression
S. No. CCTV video Compressed Encryption time per frame in sec of Decryption time per frame in sec of
video proposed algorithm proposed algorithm
highest standards of data privacy and integrity. The Kumar, A., Sharma, S., Goyal, N., Singh, A., Cheng, X.,
proposed video encryption algorithm OVEA pres- and Singh, P. (2021). Secure and energy-efficient
ents a compelling solution to the security challenges smart building architecture with emerging technol-
faced by IoT CCTV systems. Its enhanced encryption ogy IoT. Comp. Comm., 176, 207–217. [Link]
org/10.1016/[Link].2021.06.003.
times, combined with robust protection mechanisms,
Lee, D. and Park, N. (2021). Blockchain based privacy pre-
position it as a valuable addition to the arsenal of
serving multimedia intelligent video surveillance using
tools aimed at securing the interconnected world of secure Merkle tree. Multimed. Tools Appl., 80(26–27),
IoT surveillance. With the ever-changing cybersecu- 34517–34534. [Link]
rity environment, the continuous pursuit of innova- 08776-Y.
tive encryption techniques remains essential to meet Liu, F. and Koenig, H. (2005). Puzzle - A novel video en-
the evolving challenges of IoT security and maintain cryption algorithm. Lec. Notes Comp. Sci., 3677
the safety and trust of interconnected surveillance LNCS: 88–97. [Link]
systems. COVER.
Rani, S., Ahmed, S. H., and Rastogi, R. (2020). Dynamic
clustering approach based on wireless sensor net-
References works genetic algorithm for IoT applications. Wirel.
Gbashi, E. K., Shakir, E., Maolood, A. T., Gbashi, E. K., and Netw., 26(4), 2307–2316. [Link]
Mahmood, E. S. (2022). Novel lightweight video en- S11276-019-02083-7/METRICS.
cryption method based on ChaCha20 stream cipher Ravikumar, S. and Kavitha, D. (2020). RETRACTED AR-
and hybrid chaotic map. Int. J. Elec. Comp. Engg., TICLE: IoT based home monitoring system with se-
12(5), 4988–5000. [Link] cure data storage by keccak–chaotic sequence in cloud
v12i5.pp4988-5000. server. J. Amb. Intel. Hum. Comput., 12(7), 7475–
Ghimire, S. and Lee, B. (2020). A data integrity verifica- 7487. [Link]
tion method for surveillance video system. Multimed. Gera, T., Singh, J., Mehbodniya, A., Webber, J. L., Sha-
Tools Appl., 79(41–42), 30163–30185. [Link] baz, M., and Thakur, D. (2021). Dominant feature
org/10.1007/S11042-020-09482-5/METRICS. selection and machine learning-based hybrid ap-
Obaida, Tameem Hameed, Abeer Salim Jamil, and Nidaa proach to analyze android ransomware. Sec. Comm.
Flaih Hassan. (2022). A Review: Video Encryption Netw., 2021, 1–22. [Link]
Techniques, Advantages And Disadvantages. Webol- 7035233.
ogy (ISSN: 1735-188X). 19(1). (2022). Abbas Fadhil Al-Husainy, Mohammed, and Bassam Al-
Hamza, A. and Kumar, B. (2020). A review paper Shargabi. (2020). Secure and lightweight encryption
on DES, AES, RSA encryption standards. Proc. model for IoT surveillance camera. International Jour-
2020 9th Int. Conf. Sys. Model. Adv. Res. Tren. nal of Advanced Trends in Computer Science and En-
SMART 2020, 333–338. [Link] gineering. 9(2): 1840–1847.
SMART50582.2020.933680. Sultana, S. F. and Shubhangi, D. C. (2017). Video encryption
Heron, S. (2009). Advanced encryption standard (AES). algorithm and key management using perfect shuffle.
Netw. Sec., 2009(12), 8–12. [Link] Int. J. Engg. Res. Appl., 07(07), 01–05. [Link]
S1353-4858(10)70006-4. org/10.9790/9622-0707030105.
Kaur, K. and Gandhi, V. (2022). Internet of Things: A study Priya, S. M., Diana Josephine, D., and Abinaya, P. (2021).
on protocols, security challenges and healthcare ap- IoT based smart and secure surveillance system us-
plications. 2022 2nd Int. Conf. Adv. Comput. Innov. ing video summarization. Lec. Notes Elec. Engg., 735
Technol. Engg., ICACITE 2022, 1206–1210. https:// LNEE: 423–435. [Link]
[Link]/10.1109/ICACITE53722.2022.9823422. 33-6977-1_32/COVER.
Kim, J., Lee, D., and Park, N. (2020). CCTV-RFID enabled Yun, J. and Kim, M. (2020). JLVEA: Lightweight real-
multifactor authentication model for secure differen- time video stream encryption algorithm for Inter-
tial level video access control. Multimed. Tools Appl., net of Things. Sensors, 20(13), 3627. [Link]
79(31–32), 23461–23481. [Link] org/10.3390/S20133627.
s11042-020-09016-z.
73 A comprehensive review of federated learning: Methods,
applications, and challenges in privacy-preserving
collaborative model training
Meenakshi Aggarwal1, Vikas Khullar2,a and Nitin Goyal3
1,2
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Department of Computer Science and Engineering, School of Engineering and Technology, Central University of
3
Abstract
Federated learning (FL) represents an advanced approach to tackling the issues linked with training machine learning (ML)
models using distributed data while upholding privacy and security. It functions by enabling collaborative model training
across a network of edge devices or servers, all without the need to transfer raw data. In place of sending data to a central
server, which could potentially compromise privacy, federated learning empowers individual devices to conduct local train-
ing on their respective data. These updates are subsequently combined to develop an enhanced global model over multiple
iteration. Additionally, as artificial intelligence (AI) becomes pervasive in novel application areas, concerns about the privacy
of data and users are on the rise. This article offers an in-depth analysis of the advancements in FL, covering a wide array of
topics including methodologies, applications, and challenges. By sidestepping the need to transfer raw data and instead fo-
cusing on sharing model updates or gradients, FL ensures the preservation of privacy and the efficient utilization of resources.
Additionally, we investigate the diverse spectrum of application domains where FL holds significance. Instances encompass
healthcare, finance, agriculture, education, Internet of Things (IoT), and industrial processes, all benefiting from the capacity
of federated learning to harness data from decentralized sources without compromising data security. This article addresses
complications such as model diversity, Non-IID (independent and identically distributed) data distribution, communication
complexities, and security vulnerabilities. Furthermore, we discuss considerations related to regulatory compliance and eth-
ics within the context of federated learning, particularly as data privacy regulations intensify.
Keywords: Federated learning, data privacy, security, computational resources, IID (independent and identically distributed)
[Link]@[Link]
a
Applied Data Science and Smart Systems 571
Related work
Federated learning is a prominent research area that
has garnered significant attention from researchers
in recent years. This focus is driven by its various
advantages, including enhanced data privacy and
reduced communication costs. In this section, we
delve into some relevant work related to federated
learning. Jawadur Rahman et al. (2021) examined
the distinctions between FL and conventional dis-
tributed machine learning. It also delved into FL’s
distinct features and challenges while also explor-
Figure 73.1 Federated learning framework
ing its present techniques and future possibilities.
The manuscript did not narrow its focus to a par-
updates all local models, and distributes this global ticular field; instead, it covered methods for address-
model for various operations by all clients. In general, ing four fundamental challenges: issues related to
traditional centralized ML approaches face challenges privacy and security. In the same manner, (Aledhari
related to computational power, training duration, et al., 2020) also presents an in-depth overview of
and, notably, security and privacy (Elbeltagi et al., related protocols and platforms, outline the chal-
2020; Ramesh et al., 2022). FL offers an operative lenges involved, and highlight real-world use cases to
solution to address data privacy and security, ensur- provide a complete understanding of FL technology.
ing that all FL participants can enjoy the benefits of AI Yang et al. (2019) present a secure approach for FL,
(Gill and Singh, 2020; Pouriyeh et al., 2022). While a encompassing horizontal FL, vertical FL, and feder-
central server concept still exists in FL, the model is ated transfer learning. This framework is designed to
trained primarily takes place locally on these devices, facilitate the exchange of data among organizations
enhancing security and privacy by minimizing data while employing FL techniques. Du et al. (2020) con-
transfer requirements. ducted a concise review of prior research regarding
As FL occurs in a distributed setting, it necessitates FL and its application within wireless Internet of
a consistent and dependable network connection Things (IoT) contexts. Following this, they delved
among end devices for continuous update sharing. into the importance and technical hurdles associated
This can be challenging because end-device network with implementing FL in vehicular IoT scenarios,
connections often have significantly lower speeds and they identified prospective avenues for future
than those found in data centers. Such communi- research in this domain. Lo et al. (2022) introduce a
cation limitations in FL can lead to potential cost set of architectural patterns aimed at addressing the
implications during the training process (Khullar and design complexities inherent in federated learning
Singh, 2022; Vimalajeewa et al., 2022). Consequently, systems. These architectural patterns offer reusable
research efforts have been dedicated to enhancing the solutions to frequently encountered issues that arise
efficiency of communication in FL environments. in the context of software architecture design. Liu
In recent years, significant research efforts have been et al. (2020) explore the challenges, methodologies,
dedicated to the field of FL, resulting in several sur- and future prospects of FL in the context of 6G com-
vey papers that have condensed insights from diverse munications. It provides a comprehensive analysis of
domains and research areas within FL. In our study, both the strengths and weaknesses associated with
we began by examining existing surveys encompass- traditional ML in the context of 6G, as well as the
ing a wide spectrum of FL research domains and focal potential for FL to enhance the feasibility of 6G com-
points. The major contributions are as follows: munications. We categorize the FL framework into
three primary domains: FL architectures, challenges
It conducts a comprehensive examination and in- inherent to FL, and application areas, as depicted in
depth analysis of recent FL survey papers. the accompanying Figure 73.2.
572 A comprehensive review of federated learning
HFL Independent, data privacy Limited to identical feature sets Collaborative prediction,
mobile apps
VFL Accommodates complementary Requires data alignment Healthcare, finance, feature
features between clients sharing
FTL Leverages pre-trained models Complexity in model Natural language processing,
coordination image recognition
DFL Enhances privacy and reduces More challenging to manage Edge computing, privacy-
centralization critical applications
CFL Ensures data ownership and Complex data sharing Cross-organizational data
control agreements analysis
EFL Reduces data transfer and Limited computational Real-time IoT, mobile AI
latency on edge devices resources on edges
Hybrid FL Flexibility in handling various Complexity in hybrid model Versatile data collaboration
data distribution creation scenarios
Applied Data Science and Smart Systems 573
Han, et al., 2016 Cost Compression The network was pruned, the weights were
quantified, and Huffman coding was applied
Yang et al., 2021 Heterogeneity Participation of client, FL simulation platform designed for researchers
related to client numbers
Anelli et al., 2019 Heterogeneity Participation of client, Enhanced aggregation through the assessment
related to data interacted by of individual device contributions using multiple
clients criteria
Hitaj et al., 2017 Threats GAN attacks Inferring class representative
Pyrgelis et al., 2018 Threats ML classifier Membership inference
574 A comprehensive review of federated learning
Education: FL in education enhances student privacy, giving ideas to upcoming researchers on how to make
enables personalized learning experiences, and aids in progress in the field of FL and related areas. This is
resource allocation and teacher training, ultimately crucial for helping future scholars figure out what they
improving the quality and effectiveness of education. can explore in the constantly changing world of feder-
Fachola et al. (2023), delve into the practical imple- ated learning. So, this document becomes a very use-
mentation of federated learning techniques within the ful tool for researchers, people who use these ideas in
context of learning analytics, specifically focusing on practice, and those who make the rules, all of whom
the crucial challenge of student dropout prediction. want to use FL to make better decisions with data
By applying federated learning to this problem, the while keeping that data safe and private.
study aims to harness the collective intelligence of
distributed data sources, such as multiple educational References
institutions or platforms, while safeguarding individ-
ual student privacy. Aggarwal, M., Khullar, V., and Goyal, N. (2022). Contem-
porary and futuristic intelligent technologies for rice
Finance: FL in finance enhances data privacy and leaf disease detection. 2022 10th Int. Conf. Reliab.
security by allowing organizations to collaborate on Infocom Technol. Optim. (Trends and Future Di-
model training without sharing sensitive financial rections), ICRITO 2022, 12–17. IEEE. [Link]
data. It enables the development of more accurate org/10.1109/ICRITO56286.2022.9965113.
fraud detection models, risk assessment algorithms, Aledhari, M., Razzak, R., Parizi, R. M., and Saeed, F.
and personalized financial service, while complying (2020). Federated learning: A survey on enabling
with stringent regulatory requirements. Byrd and technologies, protocols, and applications. IEEE Acc.,
Polychroniadou (2020) introduce a user-friendly 8, 140699–140725. [Link]
explanation of their privacy-conscious FL protocol, CESS.2020.3013541.
Anelli, V. W., Deldjoo, Y., Di Noia, T., and Ferrara, A. (2019).
using logistic regression on a real credit card fraud
Towards effective device-aware federated learning.
dataset, aimed at individuals without expertise in Lec. Notes Comp. Sci. (Including Subseries Lecture
the field. Open banking empowers customers to con- Notes in Artificial Intelligence and Lecture Notes in
trol their financial data, fostering a new era of data- Bioinformatics) 11946 LNAI, 477–491. [Link]
driven financial services. In the future, decentralized org/10.1007/978-3-030-35166-3_34.
data ownership facilitated by federated learning may Antico, T. M., Moreira, L. F. R., and Moreira, R. (2023).
become prevalent in the finance sector. Evaluating the potential of federated learning for
maize leaf disease prediction. Proc. Nat. Meet. Ar-
Conclusion tif. Comput. Intel. (ENIAC), 282–293. [Link]
org/10.5753/eniac.2022.227293.
The concept of FL is swiftly gaining traction in various Bonawitz, Keith, Hubert Eichner, Wolfgang Grieskamp,
aspects of contemporary life, with a focus on enhanc- Dzmitry Huba, Alex Ingerman, Vladimir Ivanov,
ing data security and management across multiple Chloe Kiddon et al. (2019). Towards federated learn-
domains. In essence, FL seeks to enable secure data ing at scale: System design. Proceedings of Machine
sharing and access while ensuring a seamless experi- Learning and Systems, 1: 374–388.
ence (Aledhari et al., 2020). It permits establishments Byrd, David, and Antigoni Polychroniadou. (2020).
to jointly train predictive models without the need to Differentially private secure multi-party compu-
tation for federated learning in financial appli-
disclose their data. Lately, FL has garnered increas-
cations. In Proceedings of the First ACM Interna-
ing attention in both academic and industrial circles, tional Conference on AI in Finance, 1–9. [Link]
offering solutions to data collection and privacy hur- org/10.1145/3383455.3422562
dles, particularly in sectors like healthcare, agriculture, Du, Z., Wu, C., Yoshinaga, T., Yau, K. L. A., Ji, Y., and Li,
finance, and education (Jawadur Rahman et al., 2021). J. (2020). Federated learning for vehicular Internet of
The review covers important aspects of FL, including Things: Recent advances and open issues. IEEE Open
its basic ideas, the advanced technologies behind it, the J. Comp. Soc., 1(1), 45–61. [Link]
complex structure it uses, the various problems it tack- OJCS.2020.2992630.
les, and the different ways it’s used in areas like health- Dun, Chen, Mirian Hipolito, Chris Jermaine, Dimitrios Dimi-
care, farming, finance, and education. This in-depth triadis, and Anastasios Kyrillidis. (2023). Efficient and
analysis helps us get a complete picture of FL, includ- Light-Weight Federated Learning via Asynchronous Dis-
tributed Dropout. In International Conference on Artifi-
ing how it works, what it can do, and what it can’t do.
cial Intelligence and Statistics, 206, 6630–6660. PMLR.
It’s like shining a light on all the details of this innova- Durrant, A., Markovic, M., Matthews, D., May, D., Enright,
tive approach to data analysis and sharing. Moreover, J., and Leontidis, G. (2022). The role of cross-silo fed-
this study doesn’t only focus on what has already erated learning in facilitating data sharing in the agri-
happened; it also considers what might happen in the food sector. Comp. Elec. Agricul., 193, 1–23. https://
future. It points out possible paths for more research, [Link]/10.1016/[Link].2021.106648.
Applied Data Science and Smart Systems 575
Elbeltagi, A., Aslam, M. R., Malik, A., Mehdinejadiani, IEEE Netw., 35(1), 234–241. [Link]
B., Srivastava, A., Bhatia, A. S., and Deng, J. (2020). MNET.011.2000263.
The impact of climate changes on the water footprint Lim, W. Y. B., Luong, N. C., Hoang, D. T., Jiao, Y., Liang, Y.
of wheat and maize production in the Nile delta, C., Yang, Q., Niyato, D., and Miao, C. (2020). Federat-
Egypt. Sci. Tot. Environ., 743, 140770. [Link] ed learning in mobile edge networks: A comprehensive
org/10.1016/[Link].2020.140770. survey. IEEE Comm. Surv. Tutor., 22(3), 2031–2063.
Fachola, C., Tornaría, A., Bermolen, P., Capdehourat, G., [Link]
Etcheverry, L., and Fariello, M. I. (2023). Federated Gill, Rupali, and Jaiteg Singh. (2020). A review of neuro-
learning for data analytics in education. Data, 8(2), marketing techniques and emotion analysis classifiers
1–16. [Link] for visual-emotion mining. In 2020 9th Internation-
Han, S., Mao, H., and Dally, W. J. (2016). Deep compres- al Conference System Modeling and Advancement
sion: Compressing deep neural networks with pruning, in Research Trends (SMART), 103–108. IEEE, doi:
trained quantization and Huffman coding. 4th Int. Conf. 10.1109/SMART50582.2020.9337074.
Learn. Repres. ICLR 2016 Conf. Track Proc., 1–14. Liu, Y., Kang, Y., Xing, C., Chen, C., and Yang, Q. (2020).
Hao, M., Li, H., Luo, X., Xu, G., Yang, H., and Liu, S. A secure federated transfer learning framework. IEEE
(2020). Efficient and privacy-enhanced federated Intel. Sys., 35(4), 70–82. [Link]
learning for industrial artificial intelligence. IEEE MIS.2020.2988525.
Trans. Indus. Inform., 16(10), 6532–6542. [Link] Liu, Y., Yuan, X., Xiong, Z., Kang, J., Wang, X., and Niyato,
org/10.1109/TII.2019.2945367. D. (2020). Federated learning for 6G communications:
Hitaj, B., Ateniese, G., and Perez-Cruz, F. (2017). Challenges, methods, and future directions. China
Deep models under the GAN: Information leak- Comm., 17(9), 105–118. [Link]
age from collaborative deep learning. Proc. ACM JCC.2020.09.009.
Conf. Comp. Comm. Sec., 603–618. [Link] Lo, S. K., Lu, Q., Zhu, L., Paik, H. Y., Xu, X., and Wang, C.
org/10.1145/3133956.3134012. (2022). Architectural patterns for the design of feder-
Huang, L., Yin, Y., Fu, Z., Zhang, S., Deng, H., and Liu, D. ated learning systems. J. Sys. Softw., 191, 1–19. https://
(2020). LoadaBoost: Loss-based AdaBoost federated [Link]/10.1016/[Link].2022.111357.
machine learning with reduced computational com- Pouriyeh, Seyedamin, Osama Shahid, Reza M. Parizi, Quan
plexity on IID and non-IID intensive care data. PLoS Z. Sheng, Gautam Srivastava, Liang Zhao, and Mo-
ONE, 15 (4), 1–16. [Link] hammad Nasajpour. (2022). Secure smart communi-
pone.0230706. cation efficiency in federated learning: Achievements
Islam, Moinul, Md Tanzim Reza, Mohammed Kaosar, and and challenges. Applied Sciences. 12(18): 8980 https://
Mohammad Zavid Parvez. (2023). Effectiveness of [Link]/10.3390/app12188980.
federated learning and CNN ensemble architectures Pyrgelis, Apostolos, Carmela Troncoso, and Emiliano De
for identifying brain tumors using MRI images. Neu- Cristofaro. (2017). Knock knock, who's there? Mem-
ral Processing Letters, 55(4): 3779–3809. bership inference on aggregate location data. arXiv
Jawadur Rahman, K. M., Ahmed, F., Akhter, N., Hasan, preprint arXiv:1708.06145. [Link]
M., Amin, R., Aziz, K. E., Muzahidul Islam, A. K. M., arXiv.1708.06145.
Mukta, Md S. H., and Najmul Islam, A. K. M. (2021). Ramesh, T. R., Lilhore, U. K., Poongodi, M., Simaiya, S.,
Challenges, applications and design aspects of federat- Kaur, A., and Hamdi, M. (2022). Predictive analysis of
ed learning: A survey. IEEE Acc., 9, 124682–124700. heart diseases with machine learning approaches. Ma-
[Link] laysian J. Comp. Sci., 2022(Special issue 1), 132–148.
Jiménez-Sánchez, A., Tardy, M., Ballester, M. A. G., Mateus, [Link]
D., and Piella, G. (2023). Memory-aware curriculum Vimalajeewa, D., Kulatunga, C., Berry, D. P., and Balasub-
federated learning for breast cancer classification. ramaniam, S. (2022). A service-based joint model used
Comp. Methods Prog. Biomed., 229, 1–15. [Link] for distributed learning: Application for smart agricul-
org/10.1016/[Link].2022.107318. ture. IEEE Trans. Emerg. Top. Comput., 10(2), 838–
Khullar, V. and Singh, H. P. (2022). Privacy protected in- 854. [Link]
ternet of unmanned aerial vehicles for disastrous site Yang, C., Wang, Q., Xu, M., Chen, Z., Bian, K., Liu, Y.,
identification. Concurr. Comput. Prac. Exp., 34(19), and Liu, X. (2021). Characterizing impacts of het-
1–10. [Link] erogeneity in federated learning upon large-scale
Khullar, V. and Singh, H. P. (2023). F-FNC: Privacy con- smartphone data. Web Conf. 2021 Proc. World
cerned efficient federated approach for fake news Wide Web Conf. WWW 2021, 935–946. [Link]
classification. Inform. Sci., 639, 1–15. [Link] org/10.1145/3442381.3449851.
org/10.1016/[Link].2023.119017. Yang, Qiang, Yang Liu, Tianjian Chen, and Yongxin Tong.
Li, P., Li, J., Huang, Z., Li, T., Gao, C. Z., Yiu, S. M., and (2019). Federated machine learning: Concept and
Chen, K. (2017). Multi-key privacy-preserving deep applications. ACM Transactions on Intelligent Sys-
learning in cloud computing. Fut. Gen. Comp. Sys., 74, tems and Technology (TIST). 10(2): 1–19 [Link]
76–85. [Link] org/10.1145/3298981.
Li, Y., Chen, C., Liu, N., Huang, H., Zheng, Z., and Yan, Zhu, H., Xu, J., Liu, S., and Jin, Y. (2021). Federated learning
Q. (2021). A blockchain-based decentralized feder- on non-IID data: A survey. Neurocomput., 465, 371–
ated learning framework with committee consensus. 390. [Link]
74 Review of techniques for diagnosis of Meibomian gland
dysfunction using IR images
Deepika Sood, Anshu Singlaa and Sushil Narang
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
The main factor in the health of ocular surface is the secretion of lipids by the meibomian glands into tears, where they form
a polar lipid layer that prevents aqueous evaporation. Nowadays, the increase in usage of digital screens in human life is one
of the leading causes of dysfunctioning of meibomian glands. This results in dry eye disease (DED). Since the prevalence of
Meibomian gland dysfunction (MGD) is increasing rapidly, it becomes imperative to find effective techniques to diagnose
MGD with minimal human intervention. Early detection of MGD will be helpful to provide medication in right time to
the patient. In this paper, different techniques applied to diagnose MGD have been surveyed thoroughly. This study will
provide the current status of research that has been carried out till date to diagnose MGD. We conclude with a literature
review, acknowledging the important contributions made to our present understanding of the MGs and MGD by a number
of researches.
Keywords: meibography, Meibomian gland, Meibomian gland dysfunction, IR images, morphological features
[Link]@[Link]..in
a
Applied Data Science and Smart Systems 577
Table 74.1 Ethnicity-based prevalence of MGD (Hassanzadeh et al., 2021)
Ethnicity-based prevalence of MGD African Arabic Caucasian East Asian Hispanic South Asian
morphology. Clinically, during dry-eye examination it spite well-intentioned advancements like the
is usually recommended to image and quantify MG oblique T-shaped probe (Wise et al., 2012).
(Modi et al., 2021; Setu et al., 2021). Meibography is b. Non-contact meibography
simple and non-invasive method to capture detailed The newest meibographic approach, non-con-
morphometric information of MG of the eye. tact meibography was presented by Arita et al.
Meibography has advantages such as measuring the (2008). A digitally everted eyelid is imaged us-
progression of the disease, monitoring it and assessing ing video camera that uses an IR charge-coupled
the effectiveness and potential of treatments device and a slit-lamp biomicroscope with an IR
Infrared meibography uses infrared light to create filter. Non-contact meibography differs from the
a detailed image of the MGs, which can be used by a contact method in that a light probe is not re-
healthcare provider to diagnose and treat conditions quired. This method solves the problems of lid
that affect the glands and the tear film. Infrared light is manipulation and discomfort of patient that are
used in infrared meibography because it can penetrate frequently experienced with contact meibogra-
through the eyelid and other tissues, allowing health- phy by doing away with the necessity for a light
care providers to see the meibomian glands without probe. In contrast of this, non-contact meibog-
the need for invasive procedures. Infrared light has a raphy asserts to be patient-centered, rapid and
longer wavelength than visible light, which makes it user-friendly than contact approaches. Non-con-
less likely to be scattered or absorbed by the tissues tact meibography also has the benefit of view-
of the eye, allowing for a clearer and more detailed ing a larger portion of the everted eyelid, which
image of the meibomian glands. Additionally, infrared requires fewer photos and take lesser time to
light is safe for use in the eye, as it does not produce combine into a wide view of the lid for analysis.
any harmful effects. The progress in techniques of meibography was
described who invented the mobile pen-shaped
Meibography techniques meibography system.
This system is an example of non-contact
There are three types of meibography techniques: con- meibography that makes use of an infrared LED
tact, non-contact meibography and lipid layer thick- that is linked to a pen-shaped camera that can be
ness measurement test which are discussed as below: held in the hand and that is proficient in taking
high-quality pictures or videos of the meibomian
a. Contact meibography glands. Consequently, the slit-lamp biomicroscope
It is a conventional procedure that was created that is traditionally used in non-contact meibogra-
in the late 1970s and involves directly applying phy is effectively unnecessary with the mobile pen-
a light probe to the skin to illuminate the eyelid shaped meibography device (Wise et al., 2012).
and evert it, then imaging the results with a spe- c. Lipid layer thickness measurement
cialized camera. These systems have achieved re- Lipid layer thickness measurement is a diagnostic
markable success throughout the years; however, technique that is used to evaluate the thickness of
there are some drawbacks as well. the lipid (or oily) layer of the tear film on the sur-
• The one drawback is the requirement of an face of eye. This layer is formed by the MGs in the
expert operator who can use the equipment eyelid and it helps to keep the surface of the eye
efficiently and get high-quality images. moist and healthy. A thin or insufficient lipid layer
• Eyelids have distinct features that are rare- can cause dry eye symptoms and other problems
ly susceptible to manipulate with the light with the tear film. LLT measurement is typically
probe. A typical and time-consuming prob- performed using specialized equipment such as an
lem is partial lid eversion that requires nu- interferometer or an optical coherence tomogra-
merous photos to be captured and fused for pher, which can measure the thickness of the lipid
generating a wide image of the eyelid. layer with high accuracy. This information can be
• The probe’s heat, pressure, brightness, and used by a healthcare provider to detect and treat
sharpness could make patients feel uncom- conditions that affect the meibomian glands and
fortable. the tear film (Özcura et al., 2007).
• A recently developed “oblique T-shaped
probe” that helps to enhance image qual-
Literature review and major findings
ity and lid manipulation while decreasing
patient pain was made public in year 2007 MGD is a widespread eye disorder. In literature, sev-
in order to avoid these issues. The growth eral researchers have published a number of tech-
of non-contact meibography has consider- niques for diagnosing Meibomian gland dysfunction
ably dominated contact meibography, de- in the literature as can be seen in Table 74.2.
Applied Data Science and Smart Systems 579
Table 74.2 Different techniques for diagnosis of Meibomian gland dysfunction
1 Yu et al., 2022 1878 meibography Mask R-CNN • 21 times faster method than clini-
images cians
• Grade the MGs more accurately
• Efficiency increased but various
structural anomalies and structural
defects remains to be further re-
fined
2 Zhang et al., 2022 1620 upper eyelids Mask R-CNN • Sensitivity - 88%
and 2386 lower • Specificity - 81%
eyelids images
3 Dai et al., 2021 120 meibography A CNN with • Although precision of MGD diag-
images in a enhanced mini U-Net nosis increases but data set is very
prospective trial MGs extraction
approach small
• Reduce analysis time to help oph-
thalmologists with minimal clinical
experience
4 Khan et al., 2021 706 meibography Pearson correlation, • Improves the quantification of IR
images grading, Meiboscoring irregularities
and Bland-Altman
analysis • Locating and analyzing the MGD
dropout area.
• MGD score based quantitative
evaluation is required
5 Wang et al., 2021 1443 meibography Support vector • Mean intersection over union –
images machine classifier 63%
(SVM)
• Sensitivity – 84.4%
• Specificity – 71.7%
6 Setu et al., 2021 728 meibography U-net • Average precision – 83%
images • Recall – 81%
• F1 score – 84%, respectively
7 Prabhu et al., 2020 400 prototype CNN • Length – 47.88
images and 400 • Number of glands – 14.40
oculus images
• Gland-drop – 0.56
• Tortuosity – 1.31
• Width – 4.20
8 Xiao et al., 2021 15 meibography Image contrast • Similarity index = 0.94±0.02
images enhancement and ROI • False-negative rate = 6.43±1.98%
segmentation
• False-positive rate = 6.02±2.41%
• Technique is applicable to upper lid
MGs only
9 Celik et al., 2013 131 meibography Support vector • Accuracy of MGD – 88%
images machine classifier • Focus on classification only
(SVM)
• MGD score based quantitative
evaluation is required
A detailed literature study regarding the diagno- • DED is more common in different parts of the
sis of Meibomian gland dysfunction and DED in IR world with a range of 3.9–21.8%; it has been
images has resulted in the following findings which noted that females are more prone than males
are yet to be addressed: to have the condition and older people are more
580 Review of techniques for diagnosis of Meibomian gland dysfunction using IR images
prone to develop it than younger people. Diag- required. The major drawback of this work is the
nosis is more challenging because there is no absence of analysis from the lower lid. It takes
connection between the symptoms and signs of into consideration only the upper lid analysis.
dry eye disease. This is supported by the finding Meibography images of the lower lid are usually
that 22% of individuals with MGD, the major distorted and only partially display the meibo-
cause of DED, are unaware that they have the mian glands because averting the lower eyelid is
condition. When wearing contact lenses, the co- harder than averting the upper eyelid, which has a
existence of DED presents a significant challenge larger tarsal plate. As a result, it is still difficult to
(Markoulli, 2017). The primary reason of evapo- automatically segment the region of interest and
rative dry eye disease is MGD. Managing MGD MGs in the lower lid. Therefore, measuring mei-
is a crucial component of managing the challenge bomian glands in the human eye requires image
of CLD because dryness is one of the main causes segmentation technique. Therefore, since exam-
of contact lens dropout. ining both lids simultaneously improves clinical
• There are no or only a few segmentation tech- diagnostic performance, lower lid meibography
niques available on IR images. Those present are image analysis will be the focus of future study
not effective enough to provide accurate detec- (Zhang et al., 2022).
tion of MGD. However, finding a single test that • The author Prabhu et al. (2020) introduced an
is repeatable, reliable, and widely acknowledged approach rely on DL for automatically segment-
can be utilized in general practice is still difficult. ing meibomian glands. This study was analyzed
This demand has prompted numerous academics using five significant metrics and it was discov-
and firms to create a variety of complex diagnos- ered that they accurately reflect the MGD-related
tic tools that can be modified for MGD screening alterations. The results of comparing the images
(Markoulli, 2017). As a result, it’s significant to captured by the Bosch hand-held imager with
offer exact and reliable evaluation tests to iden- those taken by the Oculus Keratograph reveal
tify the disease as soon as possible so that the that the particular images are equivalent for the
patient can receive the best possible management evaluating MGD and that the algorithm created
and treatment plan (Markoulli, 2017). by Bosch is a useful and accurate method for the
• Several authors have worked on techniques for MGD analysis. Deep learning and automatically
segmentation on infrared images for their effec- segmenting glands can simplify MGD manage-
tiveness and improved performance. A grading ment. The future of medical diagnostics is the use
system for detecting MG area in non-contact of AI and neural networks (Prabhu et al., 2020).
meibography images of the lower and upper A related study quantifies MG abnormalities like
eyelids is described in a related work. Despite as light reflection, inappropriate light focus and
efforts, the requirement for manual interven- placement and eyelid eversion. It’s quite difficult
tion cannot be totally removed from the images, to eliminate these unexpected flaws by enhanc-
manual rectification was required after the auto- ing the system that automatically detects these
matic MG recognition that was used in images reflections. In a study by Khan et al. (2021), pro-
with high meibomian gland loss and reflected posed an adversarial learning-based automatic
light. method for the precise segmentation, detection
and analysis of MGs. This technique is free from
As a result, this system is not entirely automated. the constraint of previous assessment techniques
To make this system fully automatic, more changes It makes it possible for clinics to determine the
are required. MG dropout area in a more precise manner. It
The diversity of results in particularly the terminal supports only the characterization of MG and
portion of glands also requires more investigation minimizes the associated time with the MG anal-
(Arita et al., 2014) ysis (Kanika, 2019; Khan et al., 2021). Still, an
approach is required to predict MGD score for
• The study (Xiao et al., 2021) includes additional an eye that will lead to quantitative analysis and
morphological and functional characteristics for more accurate detection of MGD.
the investigation of meibomian glands, and the
early findings indicate its potential for detection Discussion and conclusion
of the subtle variations present in meibomian
glands. To fully characterize the relationships be- The study emerging out from the literature survey
tween the quantitative measures and the clinical states that there are several limitations to the use of
expression of MGs in various pathological phases infrared meibography in the diagnosis of MGD. Some
of disease; however, large-scale clinical research is of these limitations include:
Applied Data Science and Smart Systems 581
i. The test requires specialized equipment and Arita, R., Itoh, K., Inoue, K., and Amano, S. (2008). Non-
trained personnel to perform, which can make it contact infrared meibography to document age-re-
difficult to access in some settings. lated changes of the Meibomian glands in a normal
ii. The test can be uncomfortable for some patients, population. Ophthalmol., 115(5), 911–915. https://
[Link]/10.1016/[Link].2007.06.031.
as it involves the use of a bright light and close
Özcura, F., Aydin, S., and Helvaci, M. R. (2007). Ocular
proximity to the eye. surface disease index for the diagnosis of dry eye
iii. The test can only provide a snapshot of the Mei- syndrome. Ocul. Immunol. Inflam., 15(5), 389–393.
bomian glands at a single point in time, so it may [Link]
not be able to detect changes in gland function Yu, Y., Zhou, Y., Tian, M., Zhou, Y., Tan, Y., Wu, L., Zheng,
over time. H., and Yang, Y. (2022). Automatic identification of
iv. The test may not be able to detect early stages meibomian gland dysfunction with meibography imag-
of Meibomian gland dysfunction, as the glands es using deep learning. Int. Ophthalmol., 2022, 1–16.
may not show significant changes until the con- Zhang, Zuhui, Xiaolei Lin, Xinxin Yu, Yana Fu, Xiaoyu
Chen, Weihua Yang, and Qi Dai. (2022). Meibomian
dition has progressed.
gland density: An effective evaluation index of meibo-
mian gland dysfunction based on deep learning and
In nutshell, researchers have provided the summary transfer learning. Journal of Clinical Medicine, 11(9):
of techniques which have been utilized to diagnose 2396. [Link]
MGD. Even if the literature contains efficient state-of- Dai, Q., Liu, X., Lin, X., Fu, Y., Chen, C., Yu, X., Zhang,
art techniques, still there is room to do further work Z., et al. (2021). A novel meibomian gland morphol-
in this field that may help to yield effective techniques ogy analytic system based on a convolutional neural
by using IR images to diagnose MGD with minimal network. IEEE Acc., 9, 23083–23094. [Link]
intervention at an early stage and to overcome the org/10.1109/ACCESS.2021.3056234.
above mentioned limitations. Khan, Zakir Khan, Arif Iqbal Umar, Syed Hamad Shirazi,
Asad Rasheed, Abdul Qadir, and Sarah Gul. (2021).
Image based analysis of meibomian gland dysfunction
References using conditional generative adversarial neural net-
Lam, P. Y., Co Shih, K., Fong, P. Y., Chan, T. C. Y., Ki Ng, work. BMJ Open Ophthalmology. 6(1). doi: 10.1136/
A. L., Jhanji, V., and Tong, L. (2020). A review on bmjophth-2020-000436
evidence-based treatments for Meibomian gland dys- Prabhu, S. M., Chakiat, A., Shashank, S., and Poojita, K.
function. Eye Contact Lens, 46(1), 3–16. [Link] (2020). Biomedical signal processing and control deep
org/10.1097/ICL.0000000000000680. learning segmentation and quantification of meibo-
Hassanzadeh, S., Varmaghani, M., Zarei-Ghanavati, S., mian glands. Biomed. Sig. Proc. Con. 57, 101776.
Shandiz, J. H., and Khorasani, A. A.. (2021). Global [Link]
prevalence of Meibomian gland dysfunction: A sys- Modi, Nandini, and Jaiteg Singh. (2021). A review of various
tematic review and meta-analysis. Ocul. Immunol. state of art eye gaze estimation techniques. Advances
Inflam., 29(1), 66–75. [Link] in Computational Intelligence and Communication
948.2020.17554. Technology: Proceedings of CICT 2019: 1086, 501–
Setu, A. K., Horstmann, J., Schmidt, S., Stern, M. E., and 510. [Link]
Steven, P. (2021). Deep learning ‑ Based automatic Xiao, Peng, Zhongzhou Luo, Yuqing Deng, Gengyuan Wang,
Meibomian gland segmentation and morphology and Jin Yuan. (2021). An automated and multipara-
assessment in infrared meibography. Scientif. Rep., metric algorithm for objective analysis of meibography
1–11. [Link] images. Quantitative imaging in medicine and surgery,
Mahajan, P. and Kaushal, J. (2020). Epidemic trend of 11(4): 1586–1599. doi: 10.21037/qims-20-611
COVID-19 transmission in India during lockdown-1 Celik, T., Lee, H. K., Petznick, A., and Tong, L. (2013). Bio-
phase. J. Comm. Health, 45(6), 1291–1300. https:// image informatics approach to automated meibomian
[Link]/10.1007/s10900-020-00863-3. gland analysis in infrared images of meibography. J.
Ramesh, T. R., Lilhore, U. K., Poongodi, M., Simaiya, S., Optomet., 6(4), 194–204. [Link]
Kaur, A., and Hamdi, M. (2022). Predictive analysis of optom.2013.09.001.
heart diseases with machine learning approaches. Ma- Markoulli, Maria, and Sailesh Kolanu. (2017). Contact lens
laysian J. Comp. Sci., 2022(Special Issue 1), 132–148. wear and dry eyes: challenges and solutions. Clinical
[Link] optometry: 41–48.
Wang, J., Li, S., Yeh, T. N., Chakraborty, R., Graham, A. Arita, Reiko, Jun Suehiro, Tsuyoshi Haraguchi, Rika Shi-
D., Yu, S. X., and Lin, M. C. (2021). Quantifying rakawa, Hideaki Tokoro, and Shiro Amano. (2013).
Meibomian gland morphology using artificial intel- Objective image analysis of the meibomian gland area.
ligence. 98(9), 1094–1103. [Link] British Journal of Ophthalmology. 746–55 https://
OPX.0000000000001767. [Link]/10.1136/bjophthalmol-2012-303014
Wise, R. J., Sobel, R. K., and Allen, R. C. (2012). Mei- Kanika. (2019). KelDec: A recommendation system for ex-
bography: A review of techniques and technologies. tending classroom learning with visual environmental
Saudi J. Ophthalmol., 26(4), 349–356. [Link] cues. ACM Int. Conf. Proc. Ser., 99–103. [Link]
org/10.1016/[Link].2012.08.007. org/10.1145/3342827.3342849.
75 The impact of unstable symmetries on software engineering
Lalit Sharma1, Surbhi Bhati2, Mudita Uppal3 and Deepali Gupta4,a
1
Jaipuria Institute of Business, Ghaziabad, Uttar Pradesh, India
2
Assistant Manager, Radio city 91.1 FM
3,4
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
Abstract
The examination of Markov models is a fundamental issue, and given the present state of widespread configurations, pro-
fessionals in cyber informatics have a prudent inclination towards evaluating telephony, which involves the practical prin-
ciples of machine learning (ML). The primary objective of this study does not pertain to the recursive enumerability of the
foundational ubiquitous algorithm utilized in the enhancement of simulated annealing. Instead, the focus lies in presenting
a comprehensive framework for voice-over-IP. The primary research contribution of this study is to empirically validate
three hypotheses pertaining to the characteristics of scheme’s signal-to-noise ratio, ROM throughput, and flash memory
performance. The authors express gratitude for the utilization of wide-area networks, as they contribute to the optimiza-
tion of complexity. However, they overlook the measurement of effective reaction time and instead place emphasis on the
significance of optical drive speed. This paper elucidates the process of packet deployment at the network level, as well as
the accompanying network modifications and software implementation. Additionally, it provides a comprehensive analysis
of the hardware and software configurations involved in this process. The validation of implementation efforts is conducted
by the execution of innovative experiments that investigate factors such as energy efficiency, block size, and sampling rate.
Keywords: unstable symmetries, Markov models, software engineering, pervasive configurations, voice-over-IP
[Link]@[Link]
a
Applied Data Science and Smart Systems 583
Panda, 2021). Data mining is a commonly employed of sensor networks. However, their work did not
technique in order to derive conclusions from data extensively address the implications of web browsers
and facilitate the process of decision-making. This at the same time period (Johnson, 2005). In a subse-
paper provides a comprehensive evaluation of several quent study, Antony et al. (Hartmanis and Watanabe,
data mining and machine learning methods, encom- 1992) and Van Jacobson (Tarjan and Davis, 1994;
passing an examination of their respective algorithms, Minsky, 2000; Darwin, 2004) provided the initial
benefits, and limitations. The system aids users in the impetus for the introduction of random algorithms
selection of optimal tools for enhancing decision-mak- (Cook et al., 2003; Darwin, 2004; Singh et al., 2020).
ing, in accordance with their specified requirements Clearly, despite substantial endeavors in this field,
(Verma et al., 2019; Uppal et al., 2022). The imple- the aforementioned methodology is unquestionably
mentation of novel technologies in software engi- the favored algorithm among physicists (Darwin,
neering expedites the development process, resulting 2004; Sato et al., 2003). This technique demonstrates
in time and cost savings, as well as enhanced quality. greater robustness in comparison to our own.
This research investigates the potential of technology
to enhance the efficiency of software engineering pro-
Psychoacoustic configurations and
cesses by addressing challenges associated with dif-
implementation
ferent phases. The document includes dedicated parts
that provide an exposition on the subjects of Software The research conforms to a prescribed set of principles.
Engineering and Artificial Intelligence, examine The approach being examined by the authors entails
developing technologies, discuss the role of Artificial the aggregation of a set of n 802.11 mesh networks.
Intelligence within Software Engineering, and con- The accuracy of this assertion is uncertain in real-
clude the discussion with references to sources (Uppal world situations. Expanding on this line of argumen-
and Gupta, 2020; Uppal et al., 2022). However, they tation, the Bolye model incorporates four discrete and
faced administrative challenges that hindered their independent components, specifically event-driven
ability to share their findings until now. All of these epistemologies, Byzantine fault tolerance, homog-
strategies question the fundamental assumption that enous epistemologies, and the partition table. Instead
the use of simulated annealing and “fuzzy” informa- of formulating psychoacoustic data, Bolye chooses to
tion is characterized by a lack of clarity and confusion. incorporate lossless information. In contrast to the
predominant viewpoints held by scholars, it is impera-
Telephony tive for the algorithm to depend on this specific prop-
This approach pertains to the investigation of auton- erty in order to demonstrate precise functionality. The
omous methodologies, electronic methodologies, and initial methodology devised by Robin Milner shares
object-oriented languages. The selection of subject certain resemblances with the model currently under
matter experts in the study referenced as (Hartmanis examination; nonetheless, it successfully accomplishes
and Watanabe, 1992) diverges from our approach in the desired goal. This remark seems to be applicable
that the authors solely focus on the development of in most cases. The heuristic utilized in this study con-
confirmed archetypes within the application, as stated sists of four separate elements, specifically multi-pro-
in reference Ramasubramanian (1990). Boyle is cessors, evolutionary programming, digital-to-analog
closely linked to the study undertaken by Zhou et al. converters, and wide-area networks. While it is not
in the field of electrical engineering (Lampson et al., typical for electrical engineers to produce precise cal-
2003). Nevertheless, the authors adopt a unique per- culations in the other direction, this heuristic depends
spective in addressing this topic, placing emphasis on on this attribute to demonstrate accurate behavior.
the notion of trainable symmetries (Singh et al., 2019; To conduct a comprehensive examination of the
Corbato et al., 2002; Inder et al., 2020). In this study, location-identity split phenomena, it is crucial to
the writers have thoroughly examined and discussed verify the soundness of the decentralized algorithm
the various challenges and limitations that were pres- put out by H. Jackson in addressing the producer-
ent in the prior research. The authors intend to incor- consumer problem (Johnson, 2005). This verification
porate numerous concepts from the aforementioned process should ascertain the algorithm’s possession
previous study into future iterations of Bolye. of Turing completeness. This condition is equally
applicable to Boyle. The accuracy of this assertion
Multicast systems is uncertain in real-world scenarios. Furthermore,
This approach pertains to the study of verifiable epis- rather than imposing control on expandable modali-
temologies, the complex integration of RPCs and ties, the system chooses to explore event-driven tech-
electronic commerce, and the utilization of Lamport niques (Ullman and Sharma, 1999). Therefore, the
clocks. Robinson and Williams (Yao and Codd, 2004) methodology utilized in this study is firmly grounded
introduced a theoretical framework for the analysis in empirical facts.
584 The impact of unstable symmetries on software engineering
The authors present the latest iteration of Bolye, of efficient response time. The authors want to elu-
denoted as version 4.2.5, which has undergone a cidate the significance of enhancing the operational
thorough design process. The equitable assignment efficiency of atomic information’s optical drive speed
of permissions to both the centralized logging facility in relation to performance analysis.
and the server daemon is of paramount significance.
The construction of the hand-optimized compiler was Hardware and software configuration
rather uncomplicated when contrasted with the inher- Figure 75.2 illustrates the correlation between the
ent impracticality of the framework. In addition, the median energy of Bolye and the hit ratio. To fully
hand-optimized compiler includes about 69 instances grasp the causes of the consequences, it is crucial
of semi-colons in Perl. The decision to place a limita- to possess a thorough comprehension of the net-
tion on the block size employed by Boyle to 98 bytes work configuration. The researchers performed a
was deemed to be of utmost importance. The need packet-level implementation on the PlanetLab clus-
to impose a restriction on the usage of e-commerce ter to question the assumption that knowledge-based
by Bolye was acknowledged, resulting in the adoption archetypes possess intrinsic resilience against exter-
of a maximum limit of 9792 connections per second. nal effects. The authors had challenges in obtain-
The visual representation of the newly created opti- ing the necessary RISC processors. At the outset,
mum models is illustrated in Figure 75.1. the researchers made the decision to exclude some
central processing units (CPUs) from the metamor-
Results and discussion phic cluster. Furthermore, the researchers integrated
a Wi-Fi throughput of 200 MB/s into the system to
The evaluation is now being discussed by the authors. enhance their understanding of the network in a more
The primary objective of the overall evaluation is to
substantiate three hypotheses: (1) the basic dissimilar-
ity in the behavior of flash memory speed on the sys-
tem, (2) the fundamental dissimilarity in the behavior
of ROM throughput on the system; and lastly, (3) the
observed duplication of the signal-to-noise ratio over
time in the context of scheme. The authors express
their gratitude for the availability of replicated wide-
area networks, as these networks are essential for
their ability to optimize for complexity while adher-
ing to complexity restrictions. In contrast to their
counterparts, the authors have made the deliber-
ate choice to exclude the measurement of effective
response time. In contrast to their counterparts, the
authors have deliberately omitted the examination Figure 75.1 Illustration of the new optimal models
Figure 75.2 Relationship between the median energy of Bolye and the hit ratio
Applied Data Science and Smart Systems 585
Figure 75.3 Relationship between the average block size of Bolye and the popularity of reinforcement learning
Figure 75.4 Median sampling rate comparison of Bolye with other approaches
thorough manner. In addition, the researchers were and the degree of prominence observed within the
able to effectively reduce the bandwidth of mobile domain of reinforcement learning.
telecommunication devices. In a similar manner, the The establishment of a suitable software environ-
researchers improved the speed of the USB key device ment necessitated a significant investment of effort;
with the purpose of examining the desktop worksta- nonetheless, the ultimate result proved to be highly
tions at the Massachusetts Institute of Technology. advantageous. The authors developed the write-ahead
The configuration of this phase was found to be a logging server using PHP, incorporating jointly sto-
labor-intensive undertaking; yet, the resultant out- chastic extensions. The software was developed using
come was considered to be of significant worth. In a manual process utilizing a traditional toolchain,
order to test the speed of random access memory with the inclusion of B. Varun’s libraries. Its primary
(RAM) in the retired Motorola bag telephones from objective was to conduct computational analysis
UC Berkeley, the researchers decided to integrate on interconnected 5.25” floppy disks. Similarly, the
Non-Volatile Random-Access Memory (NV-RAM) authors have ensured that all of the software is acces-
into the system. Had the authors opted to deploy sible under the Old Plan 9 License.
the autonomous cluster in an uncontrolled environ-
ment as opposed to a controlled one, they would have Dogfooding KamMone
noticed diminished outcomes. Figure 75.3 depicts the The authors offer a justification for the substan-
relationship between the average block size of Bolye tial endeavors they dedicated to the process of
586 The impact of unstable symmetries on software engineering
Figure 75.5 (a) Relationship between interrupt rate growth and decreasing popularity of link-level acknowledgments
Figure 75.5 (b) Average energy of the approach in relation to instruction rate
implementation. Nevertheless, this claim remains discontinuities shown in the graphs indicates a reduced
valid just within the context of theoretical discourse. level of effective popularity for operating systems that
Utilizing this advantageous configuration, the team are introduced alongside hardware modifications.
executed four pioneering experiments. The research- Moreover, it is important to acknowledge that infor-
ers ran a series of tests utilizing 96 distributed nodes mation retrieval systems demonstrate more uniform
inside the Internet-2 network to execute symmet- consumption patterns of hard disk space in com-
ric multiprocessing tasks. They then proceeded to parison to exokernelized public-private key pairs.
assess the performance of these distributed execu- The findings presented in this study regarding block
tions against locally operating neural networks. The size exhibit disparities when compared to the results
researchers strategically distributed 29 UNIVACs over reported in prior studies, as shown by the influential
the Internet-2 network and subsequently performed investigation conducted by Scott Shenker on massively
kernel testing. The researchers distributed a total of multiplayer online role-playing games and the seeming
40 Nintendo Gameboys around a vast network and ubiquity of sensor networks (Lampson et al., 2003).
conducted experiments on the corresponding Web The researchers have noted a specific behavior in
services. The researchers conducted a quantitative Figure 75.5 (a), while the other tests demonstrate
analysis to determine the correlation between tape divergent results. Figure 75.5 (a) depicts the observed
drive capacity and floppy disk data transfer rate on a and unforeseen fluctuations in the speed of the effec-
Macintosh SE computer. tive RAM. Moreover, the data depicted in Figure 75.5
The authors commence their investigation by con- (b) offers persuasive substantiation that the project,
ducting a thorough analysis of all four research, as which included a period of four years, failed to yield
depicted in Figure 75.4. The existence of several the intended results, hence deeming the endeavors
Applied Data Science and Smart Systems 587
invested in it as unfruitful. The authors have con- Cook, S., Davis, V., Takahashi, K., and Johnson, M. (2003).
strained anticipations regarding the degree of preci- Synthesis of evolutionary programming. Proc. PODS.
sion attained in this specific phase of the evaluation 8, 58–62.
process. The aim of this endeavor is to offer precise Corbato, F., Bachman, C., and Sasaki, B. (2002). Compar-
ing telephony and digital-to-analog converters. Proc.
and verifiable information.
Workshop Event-Driven Encryp. Theor.
The authors proceed to discuss the experiments
Darwin, C. (2004). Deconstructing the UNIVAC computer
denoted as (3) and (4) in the previously indicated with Besee. Proc. USENIX Tech. Conf. 12, 122–133.
enumeration. The cumulative distribution function Hartmanis, J. and Watanabe, F. (1992). Simulating replica-
depicted in Figure 75.5 (a) exhibits a distinct heavy tion using authenticated archetypes. OSR, 37, 1–14.
tail, which suggests the existence of duplicated band- Inder, S., Aggarwal, A., Gupta, S., Gupta, S., and Rastogi,
width. In addition, it is crucial to note the prominent S. (2020). An integrated model of financial literacy
tail exhibited by the cumulative distribution function among B–school graduates using fuzzy AHP and fac-
(CDF) illustrated in Figure 75.2, indicating a height- tor analysis. The Journal of Wealth Management.
ened level of throughput. The curve depicted in Figure Iverson, K., Daubechies, I., and Iverson, K. (2004). On the
75.5 (b) can be identified as F*(n), a function that is evaluation of online algorithms. Proc. USENIX Sec.
Conf.
frequently denoted as (n + logn).
Johnson, D. (2005). Deconstructing IPv4 with STUFF. Proc.
Workshop Data Min. Knowl. Discov.
Conclusion Lampson, B., Moore, Z., Clarke, E., and Clarke, E. (2003).
An improvement of erasure coding with TRABU. J.
Boyle is poised to surmount a number of substantial Game-Theor. Large-Scale Introspec. Config., 75, 1–19.
challenges faced by analysts in the present period. The Minsky, M. (2000) . Controlling the memory bus and evo-
achievement of the job is dependent on the paramount lutionary programming. Proc. MOBICOM. 37, 303–
significance of this aspect. The authors conducted a 311.
comprehensive evaluation to substantiate their three Ramasubramanian, V. (1990) . A case for hash tables. Proc.
primary expectations. Firstly, they observed that flash ECOOP. 11, 265–272.
memory speed exhibits distinct behavior on the sys- Robinson, S. and ErdÖS, P. (2001) . Autonomous informa-
tem. Secondly, they noted that ROM throughput also tion for context-free grammar. Proc. NDSS. 13, 78–85.
displays different behavior. Lastly, they observed that Sato, A., Simon, H., and Kahan, W. (2003) . An investiga-
tion of wide-area networks. Tech. Rep. Microsoft Res.,
scheme consistently exhibits increased signal-to-noise
2, 166–178.
ratio over time. Wide-area networks were employed
Singh, J., Goyal, G., and Gupta, S. (2019). FADU-EV an
with the aim of optimizing complexity, deliberately automated framework for pre-release emotive analy-
excluding the evaluation of effective response time. sis of theatrical trailers. Multimed. Tools Appl., 78,
Following systematic adjustments to the hardware 7207–7224.
and software settings, experiments were done to Singh, S., Singh, J., and Sehra, S. S. (2020). Genetic-inspired
observe variations in energy, block size, and sam- map matching algorithm for real-time GPS trajecto-
pling rate. The findings underscore the importance ries. Arab. J. Sci. Engg., 45(4), 2587–2603.
of efficient optical drive speed in the examination of Tarjan, R. and Davis, X. (1994). Deconstructing symmetric
performance, notwithstanding the challenges encoun- encryption using Phycomater. Proc. Workshop Inter-
tered during implementation. The review focused on ac. Config. 23, 198–206.
Ullman, J. and Sharma, L. (1999). Study of congestion
specific experimental findings and their significance in
control. Proc. Conf. Large-Scale Distrib. Inform. 5,
order to justify the thorough implementation efforts.
109–115.
One potential constraint of the system is to its capac- Uppal, M. and Gupta, D. (2020). The aspects of artificial in-
ity to retain redundant investigations, a matter that telligence in software engineering. J. Comput. Theoret.
the authors intend to investigate further in subsequent Nanosci., 17, 4635–4642.
research endeavors. The authors predict that there Uppal, M., Gupta, D., and Mehta, V. (2022). A bibliomet-
will be a notable shift of security specialists towards ric analysis of fault prediction system using machine
the utilization of mimicking Bolye in the near future. learning techniques. Challeng. Opport. Deep Learn.
Appl. Indus., 4, 109.
Verma, Kanupriya, Sahil Bhardwaj, Resham Arya, Mir Sa-
References lim Ul Islam, Megha Bhushan, Ashok Kumar, and Piy-
Anderson, N. (1999). The influence of constant-time ar- ush Samant. (2019). Latest tools for data mining and
chetypes on e-voting technology. Proc. USENIX Sec. machine learning. International Journal of Innova-
Conf. 6, 14–19. tive Technology and Exploring Engineering (IJITEE)
Badotra, S. and Panda, S. N. (2021). SNORT based early 8(9S): 18–23.
DDoS detection system using Opendaylight and open Yao, A. and Codd, E. (2004). Towards the refinement of
networking operating system in software defined net- hierarchical databases. Proc. Conf. Constant-Time
working. Clust. Comput., 24, 501–513. Encryp. Epistemol. 17, 358–366.
76 Integrating quantum computing models for enhanced
efficiency in 5G networking systems
Anand Singh Rajawat1, S. B. Goyal2,a, Jaiteg Singh3 and Xiao Shixiao4
1
School of Computer Science & Engineering, Sandip University, Nashik, Maharashtra, India
2
City University, Petaling Jaya, 46100, Malaysia
3
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
4
Chengyi College, Jimei University, Xiamen, 3611021, China
Abstract
The emergence of 5G networking systems is a notable advancement in communication technology, offering unparalleled
data speeds and connectivity. Nevertheless, the growing complexity of these systems poses a significant challenge in terms of
ensuring both efficiency and security. This study investigates the incorporation of quantum computing models as a means
to improve the effectiveness and robustness of 5G networking systems. Quantum computing presents a promising solution
for addressing the various obstacles encountered by 5G networks, owing to its capacity to execute intricate computations
swiftly and safely. Through the utilization of quantum algorithms and the application of fundamental concepts such as su-
perposition and entanglement, the objective of this integration is to optimize the management of network traffic, strengthen
the encryption of data, and improve the overall efficiency of the system. This paper presents a thorough examination of the
potential applications of quantum computing models in many domains of 5G networking, encompassing data transmission,
network architecture, and security protocols. Additionally, the paper addresses the various obstacles and constraints associ-
ated with the process of integrating these elements. The results indicate that quantum computing models have the potential
to improve the efficiency of 5G networks and contribute to the development of more resilient and secure communication
systems in the future.
Keywords: Quantum computing, 5G networks, network efficiency, quantum algorithms, data encryption, communication
technology
drsbgoyal@[Link]
a
Applied Data Science and Smart Systems 589
The paper is organized as in the following man- to improve the effectiveness and safeguard the integ-
ner – the related work, proposed methodology, results rity of computational resources inside power service
analysis and finally conclusion and future work. systems. By delegating computing activities to edge
devices, this methodology effectively reduces latency
Related work and optimizes resource allocation, offering a feasible
resolution for managing computational requirements
The literature review conducted on the progressions in power services within the context of 5G technol-
in network security and data protection within the ogy. Each of the aforementioned research makes
framework of 5G and 6G networks uncovers a wide valuable contributions to the continuously develop-
array of inventive methodologies. The study conducted ing domain of network security and data protection,
by Mirtskhulava et al. (2021) highlights the signifi- by focusing on distinct difficulties that have emerged
cance of ensuring the security of medical data within with the introduction of 5G and 6G technologies. The
the context of 5G and 6G networks. The researchers aforementioned statements jointly emphasize the sig-
propose the utilization of multichain blockchain tech- nificance of pioneering approaches in safeguarding
nology, complemented by post-quantum signatures, data and communications amidst the emergence of
as a means to enhance the security measures in place. new technologies and potential risks (Table 76.1).
The proposed model put out by the authors suggests
the utilization of blockchain technology as a decen-
Methodology
tralized and tamper-resistant platform for the storage
of medical data. Additionally, the model incorporates Quantum computing models integration with 5G net-
post-quantum signatures to provide long-term secu- working systems is a novel technique to improve the
rity against potential attacks from quantum com- overall performance, security, and efficiency of com-
puters. This strategy effectively acknowledges the munication networks. Through this integration, com-
pressing requirement for strong data security mea- plicated problems can be solved more quickly than
sures inside healthcare systems, specifically in light with traditional computing techniques by utilizing the
of the evolving realm of 5G and 6G technologies. special powers of quantum computing, such as quan-
In a study titled “Concealed quantum tele computa- tum entanglement and superposition. domain of 5G
tion for anonymous 6G URLLC networks” published networking systems, the incorporation of quantum
in 2023, Zaman et al., investigate the potential of computing paradigms offers a revolutionary strat-
concealed quantum tele computation as a means to egy that has the potential to greatly augment both
augment anonymity in Ultra-Reliable Low-Latency efficiency and performance. This overview examines
Communication (URLLC) networks within the the possible implications of quantum computing
context of 6G technology. This novel methodology on 5G networks (Kakaraparty et al., 2021) explor-
employs principles of quantum computing to provide ing how this advanced technology might be utilized
safe and anonymous communication, a crucial aspect to overcome current constraints and enable novel
in important domains such as military operations or functionalities.
confidential corporate communications. This paper
offers a forward-looking viewpoint on the potential The emergence of 5G networks
integration of quantum technologies into forthcom- The advent of 5G technology represents a notable
ing network topologies to bolster security measures. advancement over its predecessors, since it provides
In a recent publication by Sahoo and Samantaray enhanced data transfer rates, reduced communication
(2020) undertake a comprehensive examination of delays, and increased network capacity. This techno-
non-linear photonic-based network devices, with a logical development plays a vital role in facilitating
specific emphasis on addressing the issues associated the increasing number of interconnected devices and
with big data. The study emphasizes the potential the data-intensive applications of the contemporary
of these devices in enhancing data management and digital age, including the Internet of Things (IoT),
processing capacities within network infrastructures, smart cities, and augmented reality.
a critical aspect in the context of the big data era.
This study contributes to the comprehension of the Challenges in current 5G network
integration of optical technologies into current and Despite the significant technological developments in
prospective networks for the purpose of enhancing 5G, there exist certain issues pertaining to network
the management of extensive data volumes with more management, security, and data processing capacities.
efficiency. In their recent study, Yu et al. (2022) pres- The increasing volume and intricacy of data traffic
ent a novel solution that focuses on secure compute provide challenges for traditional computing models,
offload for power services, utilizing the capabilities of resulting in difficulties in maintaining pace, which in
5G edge computing. The objective of this approach is turn give rise to bottlenecks and security concerns.
590 Integrating quantum computing models for enhanced efficiency in 5G networking systems
Table 76.1 Literature review on quantum computing and network security
Xuan et al., 2021 Quantum computing Delays are greatly Available quantum Need for more
reduces network function reduced, improving computing resources accessible quantum
visualizations latency network efficiency are scarce and computing resources
expensive for wider use
Debbabi et al., Overview of B5G and 6G Explains AI’s role in Training AI Develop lightweight
2022 network slicing resource sophisticated network algorithms requires AI models for
management AI techniques resource management vast datasets and network resource
computing management
Sabaawi et al., Creating a restricted Improves network Current network Quantum algorithm
2022 quantum optimization performance by infrastructure integration with
technique for MIMO allocating electricity installation issues MIMO systems
power allocation more efficiently
Xu et al., 2022 Securing 5G access with Improves 5G security Real-world Simplification and
semi-random coding with quantum implementation of use of quantum-based
and quantum amplitude methods quantum amplitude security
amplification amplification is
difficult
Takalkar and Quantum cryptography Quantum The theoretical study Connecting quantum
Shiragapur, 2023 security analysis and cryptography may not address cryptography theory
mathematical modeling security is thoroughly real-world application and practice
examined issues
The term “enhancedNenhanced” refers to a 5G net- necessary number of qubits and implementing the
working system that has been improved to include appropriate quantum gates.
quantum AI technology. The equation presented Data encoding: The encoding of network data is
signifies the integration of conventional networking achieved by mapping it onto quantum states (Guo et
systems (Mehic et al., 2023) with the amalgamation al., 2022), which can subsequently be manipulated by
of AI algorithms and quantum computing models, the quantum circuit.
resulting in the emergence of quantum AI algorithms
Quantum processing: The manipulation of data
within the upgraded 5G network. The algorithms
within a quantum circuit involves the utilization
known as QAI has the ability to analyze network data
of quantum gates and entanglement, enabling the
with enhanced efficiency and accuracy, hence leading
extraction of significant insights that may elude clas-
to network management that is more intelligent and
sical algorithms.
responsive.
Classical conversion: The results obtained from the
quantum circuit are transformed into a classical rep-
Algorithm: QuantumAI_5GEnhancement resentation that is intelligible and compatible with the
Inputs: AI model.
quantum_data: Data from 5G network encoded in
AI analysis: The AI model utilizes quantum-enhanced
quantum states
data processing techniques to employ machine learn-
classical_data: Classical data from 5G network
ing (Singh et al., 2020; Wang et al., 2021) algorithms
infrastructure
for the purpose of analyzing network performance
QAI_model: Pre-trained Quantum AI model for
and detecting potential areas of enhancement.
network analysis
Procedure: Management actions: The study conducted by the
1. Initialize Quantum Processing Unit (QPU) AI yields actionable insights and recommendations
2. Encode classical_data into quantum states: aimed at improving the efficiency and security of the
For each data_point in classical_data: 5G network (Xin et al., 2020).
Convert data_point to quantum state (qubit
representation) Results analysis
Add the qubit to quantum_data
3. Apply Quantum AI Model: Data throughput optimization: This gauges the speed
Load QAI_model into QPU at which information is sent. About 10 Gbps was the
For each quantum_datapoint in quantum_data: throughput rate attained by classical AI; quantum AI
Process quantum_datapoint using QAI_model increased this rate to 15 Gbps, a 50% increase.
Measure the output state to get network insights Network latency reduction: With quantum AI,
4. Quantum-Classical Hybrid Analysis: latency the amount of time before a data transfer
Combine insights from QAI_model with classical starts was lowered by 40% from 10 milliseconds (ms)
AI algorithms with classical AI to 6 ms.
Analyze combined data for enhanced network Error rate in data transmission: The percentage of
understanding data transmission failures was evaluated in this case.
5. Network Optimization Decisions: The mistake rate for classical AI was 2%; quantum AI
Based on analyzed data, make decisions for net- dramatically decreased this to 0.5% – a 75% reduction.
work optimization Resource allocation efficiency: This measure
Adjust network parameters for improved effi- assesses how well resources are being used. The usage
ciency and performance rate of classical AI was 70% with quantum AI, this
6. Feedback Loop: was increased to 90%, a 28.6% increase.
Continuously feed network performance data Traffic prediction accuracy: With quantum AI, the
back into QAI_model accuracy of predicting network traffic increased by
Retrain or adjust QAI_model based on ongoing 11.8%, from 85% with classical AI to 95%.
performance metrics Network security threat detection: The network’s
Output: security threat detection rate was evaluated in this
Optimized network parameters and configura- case study. The detection rate rose by 8.9–98% with
tions for enhanced 5G efficiency quantum AI as opposed to 90% with classical AI
End Algorithm (Table 76.2).
Analysis summary
Quantum circuit initialization: This stage entails The amalgamation of quantum computing and
configuring the quantum circuit by allocating the AI within 5G networks demonstrates a significant
592 Integrating quantum computing models for enhanced efficiency in 5G networking systems
Table 76.2 Test scenario
enhancement across many parameters in contrast and strengthened detection of security threats. The
to conventional AI systems. The primary areas of findings of this study suggest that the integration of
improvement encompass heightened data transmis- quantum AI algorithms has the potential to greatly
sion capacity, diminished time delay, decreased occur- enhance the functionalities of 5G networks, result-
rence of errors, enhanced allocation of resources, ing in improved network management characterized
improved accuracy in predicting traffic patterns, by increased intelligence and responsiveness. The
Applied Data Science and Smart Systems 593
enhancements observed are not solely gradual, but increasingly imperative, particularly in light of the
rather possess revolutionary qualities. These advance- potential of quantum computing to undermine exist-
ments indicate a potential shift in the prevailing para- ing cryptographic standards. Notwithstanding these
digm regarding the optimization and security of 5G hurdles, the potential advantages of incorporating
networks. This shift is attributed to the incorporation quantum computing and AI into 5G networks are
of quantum computing models (Xin et al., 2020). of considerable magnitude and should not be dis-
Table 76.3 presents the analysis of results obtained regarded. The integration being discussed not only
from the implementation of quantum AI in 5G net- holds the potential for improved efficiency and per-
works. The objective of this analysis is to evaluate formance, but also serves as a foundation for future
the performance and effectiveness of quantum AI advancements in network technology. With the ongo-
in enhancing the capabilities of 5G networks. The ing progress in research and development within this
results are categorized into several metrics, including domain, it is foreseeable that a forthcoming epoch
network speed, latency, security, and energy of networking systems will emerge, characterized by
enhanced speed, efficiency, intelligence, and security.
Conclusion In summary, the incorporation of quantum com-
puting models and AI into 5G networking systems
As we approach the apex of our investigation into signifies a substantial advancement in our pursuit
the fusion of quantum computing with 5G net- of more sophisticated, effective, and secure telecom-
works, it becomes apparent that this merger signi- munications networks. This innovative methodology,
fies a significant advancement in telecommunications although in its early developmental phase, has the
technology. The integration of quantum computing, potential to fundamentally transform our under-
renowned for its exceptional computational prowess, standing and engagement with network technologies,
with the adaptive and intelligent nature of AI, engen- thereby paving the way for a dynamic future in the
ders a synergistic phenomenon that substantially field of digital communications.
amplifies the functionalities of 5G networks. The
integration of quantum AI within 5G networks facil- References
itates the use of advanced data analysis techniques
and network management strategies. Quantum AI Sahoo, A. and Samantaray, L. (2020). Nonlinear photonic
algorithms, known for their capacity to efficiently based network devices to meet big data challenges: A
review. 2020 Int. Conf. Comp. Sci. Engg. Appl. (ICC-
process extensive datasets and perform intricate
SEA), 1–4.
computations at impressive velocities, demonstrate Takalkar, A. and Shiragapur, B. (2023). Quantum cryptog-
a notable aptitude for analyzing the substantial raphy: Mathematical modelling and security analysis.
volume of data produced within 5G networks. The 2023 3rd Asian Conf. Innov. Technol. (ASIANCON),
improved analytical capability results in better accu- 01–07.
racy in forecasting and decision-making, thereby Sabaawi, A. M. A., Almasaoodi, M. R., El Gaily, S., and
optimizing the performance and efficiency of the net- Imre, S. (2022). New constrained quantum optimiza-
work. Moreover, quantum AI plays a significant role tion algorithm for power allocation in MIMO. 2022
in the advancement of intelligent and highly adaptive 45th Int. Conf. Telecomm. Sig. Proc. (TSP), 146–149.
network management systems. These systems possess Wang, C. and Rahman, A. (2022). Quantum-enabled 6G
the capability to predict network demands, adjust to wireless networks: Opportunities and challenges.
IEEE Wirel. Comm., 29(1), 58–69.
dynamic conditions in real-time, and detect possible
Wang, D., Song, B., Lin, P., Yu, F. R., Du, X., and Guizani,
security risks with unparalleled precision. The imple- M. (2021). Resource management for edge intelli-
mentation of a proactive network management strat- gence (EI)-assisted IoV using quantum-inspired rein-
egy not only enhances the overall user experience but forcement learning. IEEE Internet of Things J., 9(14),
also reinforces network security, which is of para- 12588–12600.
mount importance in the current landscape charac- Xu, D., Ren, P., and Lu, L. (2022). Semi-random coding
terized by escalating cyber threats. Nevertheless, it with quantum amplitude amplification for secure
is crucial to acknowledge that the incorporation of access authentication in future 5G communications.
quantum computing and AI into 5G networks pres- 2022 8th Int. Conf. Big Data Comput. Comm. (Big-
ents certain obstacles. Significant challenges arise Com), 120–127.
from factors such as the development of quantum Xu, D., Yu, K., and Ritcey, J. A. (2021). Cross-layer device
authentication with quantum encryption for 5G en-
hardware, algorithmic complexity, and the integra-
abled IIoT in industry 4.0. IEEE Trans. Indus. In-
tion of quantum systems with pre-existing network form., 18(9), 6368–6378.
infrastructures. Furthermore, as we embark on this Debbabi, F., Rihab, J. M. A. L., Chaari, L., Aguiar, R. L.,
emerging phase of quantum-enhanced networking, Gnichi, R., and Taleb, S. (2022). Overview of AI-based
the significance of solid security protocols becomes algorithms for network slicing resource management
594 Integrating quantum computing models for enhanced efficiency in 5G networking systems
in B5G and 6G. 2022 Int. Wirel. Comm. Mob. Com- Mehic, Miralem, Libor Michalek, Emir Dervisevic, Patrik
put. (IWCMC), 330–335. Burdiak, Matej Plakalovic, Jan Rozhon, Nerman
Xin, G., Han, J., Yin, T., Zhou, Y., Yang, J., Cheng, X., and Mahovac et al. (2023). Quantum cryptography in
Zeng, X. (2020). VPQC: A domain-specific vector 5g networks: A comprehensive overview. IEEE Com-
processor for post-quantum cryptography based on munications Surveys & Tutorials. doi: 10.1109/
RISC-V architecture. IEEE Trans. Circ. Sys. I Reg. Pa- COMST.2023.3309051
pers, 67(8), 2672–2684. Singh, J., Goyal, G., and Gill, R. (2020). Use of neuro-
Abulkasim, H., Alsuqaih, H. N., Hamdan, W. F., Hamad, metrics to choose optimal advertisement method for
S., Farouk, A., Mashatan, A., and Ghose, S. Improved omnichannel business. Enterp. Inform. Sys., 14(2),
dynamic multi-party quantum private comparison 243–265.
for next-generation mobile network. IEEE Acc., 7, Xuan, W., Zhao, Z., Fan, L., and Han, Z. (2021). Minimiz-
17917–17926. ing delay in network function visualization with quan-
Sharma, H., Sharma, G., and Kumar, N. (2022). Secre- tum computing. 2021 IEEE 18th Int. Conf. Mob. Ad
cy maximization for pico edge users in 5G back- Hoc Smart Sys. (MASS), 108–116.
haul HWNs: A quantum RL approach. 2022 IEEE He, Y., Ren, Y., Zhang, L., Chen, X., and Wu, Z. (2023).
Int. Conf. Adv. Netw. Telecomm. Sys. (ANTS), Research on grid communication encryption quantum
207–212. key encryption transmission system based on 5G com-
Kakaraparty, K., Munoz-Coreas, E., and Mahbub, I. (2021). munication terminal. 2023 IEEE Int. Conf. Control
The future of mm-wave wireless communication sys- Elec. Comp. Technol. (ICCECT), 587–590.
tems for unmanned aircraft vehicles in the era of artifi- Guo, Y., Liu, G., Ren, J., Liu, Y., Yao, L., Cao, Y., Chen, J.,
cial intelligence and quantum computing. 2021 IEEE and Zhou, Y. (2022). Physical layer security of IRS-
MetroCon, 1–8. assisted multi-layer heterogeneous networks in smart
Mirtskhulava, L., Iavich, M., Razmadze, M., and Gulua, N. grid. 2022 Int. Conf. Comput. Comm. Percep. Quan.
(2021). Securing medical data in 5G and 6G via mul- Technol. (CCPQT), 128–134.
tichain blockchain technology using post-quantum Yu, Y., Wang, W., Wu, H., Qiu, L., and Xu, Y. (2022). A
signatures. 2021 IEEE Int. Conf. Inform. Telecomm. secure computing offload approach for power services
Technol. Radio Elec. (UkrMiCo), 72–75. based on 5G edge computing. 2022 IEEE 6th Adv.
Li, M., Liu, X. F., Meng, Y., and You, Q. D. (2022). A 5G Inform. Technol. Elec. Autom. Con. Conf. (IAEAC),
NTN-RAN Implementation architecture with secu- 903–909.
rity. 2022 4th Int. Conf. Comm. Inform. Sys. Comp.
Engg. (CISCE), 42–45.
77 Micro-expressions spotting: Unveiling hidden emotions
and thoughts
Parul Malik and Jaiteg Singha
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab-140401, India
Abstract
Micro-expressions (MEs), brief and uncontrollable facial movements that last only a fraction of a second, are windows
into the innermost thoughts and emotions of others. In the past decade, MEs analysis techniques have evolved gradually
from psychological-based to computer vision-based due to the ongoing advancement of computer technology. An essential
component of MEs analysis called MEs spotting has drawn increasing attention. Recently, a few review articles on MEs
have been released, however the majority of them lacked a thorough examination of MEs spotting and instead concentrated
mostly on MEs recognition. Therefore, this review article aims to conduct an analysis of the methods, uses, and constraints of
MEs spotting, a key aspect of non-verbal communication analysis. The review starts out by exploring the psychological and
physiological bases of MEs and explaining their evolutionary significance as well as their applicability in various social cir-
cumstances. MEs spotting are used in psychology to evaluate emotional states, spot dishonesty, and guide psychotherapeutic
therapies. MEs spotting improve empathy and general awareness of emotional cues in interpersonal interactions. The article,
however, is not afraid to confront the major issues facing the discipline, such as the cultural and contextual heterogeneity of
MEs, ethical issues, and the requirement for ongoing training to maintain spotting accuracy. In summary, this review article
functions as an all-encompassing reference for individuals in the fields of research, practical application, and education who
have an interest in the diverse area of identifying malicious entities. It underscores the interdisciplinary nature of this field
and its potential to revolutionize human interaction analysis, decision-making processes, and emotional well-being across a
spectrum of applications.
Keywords: Micro-expressions spotting, non-verbal communication, spot dishonesty, decision making, contextual heterogeneity
a
[Link]@[Link]
596 Micro-expressions spotting: Unveiling hidden emotions and thoughts
Table 77.1 Few of the available published research reports on MEs analysis of recent years
Year of publication /citation Facial landmark Facial landmark Face registration Masking Face regions
detection tracking
a few presented a study on MEs interval and apex applying masking to the face, and retrieving the facial
detection. In the identification of facial malicious enti- region. The details of each process are mentioned
ties, typical pre-processing steps involve tasks such as below. Table 77.2 summarizes the present pre-pro-
detecting and tracking facial landmarks, registering cessing approaches used in facial MEs spotting.
and masking the face, and retrieving the facial region.
Facial landmark detection and tracking
Pre-processing The primary and pivotal stage in the spotting frame-
The typical pre-processing procedures in the identi- work for locating facial points in facial images is facial
fication of facial malicious entities involve detect- landmark detection, as indicated by Polikovsky et al.,
ing and tracking facial landmarks, registering and (2009). Additionally, a tracking algorithm is utilized
Applied Data Science and Smart Systems 597
to track the facial points initially identified manu- two types of methods currently used to detect facial
ally in the first frame, as outlined by Polikovsky et al. micro-movements and may convey the person’s real
(2013). Later on, different automatic facial landmark emotions. Few of the existing techniques used for
detection methods have been used for facial MEs spotting a facial ME in recent years are summarized in
spotting (Cristinacce and Cootes, 2006; Saragih et al., Table 77.3.
2009; Asthana et al., 2013; Milborrow and Nicolls,
2014; Davison et al., 2015; Wang et al., 2017; Mo et Challenges and future directions
al., 2020).
Challenges in availability of MEs dataset
Face registration The rationale for the scarcity of datasets includes
Throughout the facial MEs spotting workflow, spontaneous MEs (Takalkar et al., 2018; Zhou et al.,
area and feature-based registration methods were 2021).
used on faces to remove the excess head transla- Because of the distinctive properties of MEs, they are
tions and spins (Davison et al., 2018). Area-based difficult to identify with the naked eye. Furthermore,
and feature-based approaches are the two different eliciting MEs by emotional cues in controlled labora-
kinds of registration approaches to determine the tory settings is relatively infrequent. Moreover, label-
consistency and paired connection between both the ing MEs datasets is time-consuming and is prone
sensed and referenced images (Shreve et al., 2011; to biases from diverse annotators. This can lead to
Li et al., 2018). Another, procrustes analysis is used inaccurate labels within the dataset, which hampers
to arrange detected landmark points and establish a machine learning model training. Additionally, MEs
linear transformation among sensed and reference spotting are a very esoteric field in computer vision,
images (Xia et al., 2016). garnering less attention than prominent issues such as
object identification, categorization, face recognition,
Masking and macro-expression recognition.
In the task of spotting malicious entities in facial
expressions, the application of masking to face images Challenges in feature extraction
aims to remove noise generated by undesired facial The duration of MEs and macro-expressions is a
movements, which could impact the performance frequent criterion for identifying them. If an expres-
of the spotting task. Previously, a “T-shaped” static sion lasts longer than 0.5 seconds, it is classified as a
mask was employed to eliminate the central part of macro-expression, while shorter durations are classi-
the image consisting of eyes, nose, and mouth. This fied as MEs. For example, if a movie is taken at 30
part is not considered due to eye cascades and blink- frames per second and contains a MEs clip, there
ing, rigidness of nose and undesired large motion of will be fewer than 15 frames capturing those MEs.
mouth (Shreve et al., 2011). A binary mask is also Algorithms face a considerable issue in detecting the
employed to attain twenty FACS-based facial regions exact starting and finishing points of such transient
which are beneficial for spotting task (Davison et al., expressions in a lengthy video with a large number of
2018). frames (Oh et al., 2018).
2009/ To identify facial malicious 3D gradient K-mean Polikovsky The method can This method cannot directly
Polikovsky entities by using high-speed histogram cluster be recognized as action units measure the durations of
et al., 2009 camera and a 3D-gradient to analyze with good precision. expressions.
descriptor New dataset (Polikovsky) of
facial MEs was presented
2009/Porter To segment temporal Optical strain M Threshold USF, BU This method automatically MEs spotting in USF dataset
and ten expressions from face videos technique spots macro- and ME and were observed low. Eye
Brinke, 2008 comprises of continuous and achieved 100% accuracy in tracking algorithm may be
changing expression using spotting micro- and macro applicable to precise the
strain patterns expressions in USF and BU segment of the facial region.
dataset Computation time can also
be decreased in addition to
increased robustness
2011/Shreve To automatically spotting Optical strain Threshold USF-HD A maximum of 85% spotting
et al., 2011 facial expression in long technique accuracy was observed for
videos comprising of micro- macro-expressions and 74% of
and macro-expression using all MEs
temporal segmentation
recognize MEs by analyzing
the videos frame by frame
2013/Wang To detect and characterize 3D gradient K-mean Polikovsky This method provides finer This method was exclusively
et al., 2017 facial MEs in high-speed histogram cluster description of the timing used for evaluating staged
videos using FACS characteristics of ME facial malicious entities in
videos, and the experiment
was conducted as a
classification task, which is
not applicable for real-time
detection
2014/ Automatically detecting LBP M Threshold CASME-A This method spotted Image registration using
Polikovsky swift facial movements in technique CASME-B spontaneous MEs by affine transformation could
et al., 2013 videos through the analysis SMIC-VIS-E thresholding Chi-squared be used for pose correction
of appearance-based feature distance of LBP of center also to improve robustness in
differences frame and average of first as longer videos more adaptive
well as last frames. It also gives threshold calculation is needed
details of the spatial locations
of the movements in the facial
region
Year of Purpose Features used Movement Spotting Dataset used Conclusion Challenges and future scope
publication/ (M) / apex technique
reference (A)
2015/ Identifying subtle facial HOG M Threshold in-house In this scheme, the higher To analyze alterations
Davison et movements by employing technique dataset recall of 0.8429 and in features within both
al., 2015 Histogram of oriented F1-measure of 0.7672 were the spatial and temporal
gradients (HOG) as a achieved dimensions of a video, in
feature descriptor to addition to Chi-squared
characterize the video difference distance metrics,
sequences Earth Mover’s distance could
be effective when employed
with normalized histograms.
Furthermore, this approach
could be suitable for various
feature descriptors and
additional datasets
2015/Patel To analyze and identify the Spatio- M Threshold SMIC-VIS-E This approach attained a 95% To find AU and reduce false
et al., 2015 onset and offset frames for temporal technique area under the curve (AUC). positive rate, an approach
detected malicious entities, integration of The recorded mean onset error combining recognition and
a predetermined algorithm OF vectors was reported as 1.14 with a detection
was utilized, along with standard deviation of 3.68, could be used
determining the peak while the mean offset error
was determined to be -0.55,
accompanied by a standard
deviation of 3.52
Chitkara University Institute of Engineering and Technology, Chitkara University, Punjab, India
2
Abstract
For any country, the backbone of its economy depends on the percentage of people involved in agriculture. However, many
major agricultural regions are termed underdeveloped because of many factors, like lack or no use of modernized technolo-
gies. Seed classification is still done with the farmers’ basic knowledge, which is proven to have no mechanical validations
and hence is considered inefficient. The present research presents a mechanism for assessing the quality of rice seed that can
help provide better crop production. Machine learning (ML), a subset of artificial intelligence (AI), is used for learning the
data that is used for making predictions, recognizing patterns, making simulations in the real world, and classifying the input
data. Rice is the major source of food for a total of 80% of the population of the world. Rice is grown in different variet-
ies; hence, detecting faulty seeds and distinguishing among the varieties is another important farming aspect. This research
elaborates on a method that can be used efficiently for extracting the features of rice seeds by their identification and classifi-
cation using digital imaging. The process involved is filtering, segmentation, and edge detection as pre-processing techniques.
Keywords: Artificial intelligence, machine vision, GDP growth, resource efficiency, seed quality assessment, ANN classifier
a
ridhijindal136 @[Link]
604 Artificial intelligence and machine vision-based assessment of rice seed quality
(c) Hidden layers: Additional layers (Al-Shayea and utilizing machine and computer vision, deep learning
Bahia, 2010). (DL), and ML approaches. The system, tested with a
dataset of 9692 rice images from different varieties
Literature review grown in Punjab, achieved 93% accuracy using the
KNN classifier (Komal et al., 2022). Paddy produc-
Using computer vision for plant phenotyping is gain- tion is vital in India, with a 33% increase in GDP
ing importance, aiding in identifying plant changes. export rate in 2021.
Advances in image analysis and ML, including con- Diseases like brown spot, rice blast, sheath rot,
volutional neural networks (CNNs), have expanded sheath blight, and false smut can severely affect
its use for high-throughput phenotyping. Combining crop yield. Early detection is essential, and computer
multiple sensors allows noninvasive data acquisition, vision, specifically convolutional neural networks
enhancing our understanding of plant development (CNNs), has been employed to predict disease symp-
and responses. Automated phenotyping platforms toms. Among the four classifiers tested, Inception-V3
assist in understanding gene functions in controlled acquired the highest accuracy of 95.3% (Vignesh
conditions. For extensive field phenotyping in agricul- and Elakya, 2022). A comprehensive survey on com-
ture, remote sensing technology with image-collecting puter vision-based food grain classification methods
platforms, including unmanned vehicles, is develop- is presented, analyzing various approaches for differ-
ing. This technology will aid in predicting and pre- ent grain varieties. The review examines the process-
dicting plant traits based on phenotype/genotype ing stages in the classification pipeline, image types
relationships (Mochida et al., 2019). considered, and ground truth data generation meth-
Assessing the quality of rice is vital due to its global ods. Future challenges and needs are also discussed
consumption. Traditional manual inspection methods (Hidayat et al., 2023). Staple foods like pulses and
are labor-intensive, time-consuming, and prone to grains face issues such as adulteration and quality
errors. A real-time image processing system is intro- maintenance. Traditional methods and non-destruc-
duced to classify rice grains on the basis of their com- tive techniques are analyzed for rice starch content,
mercial value. This system automatically segments physico-chemical properties, and biochemical prop-
rice grains from the background, extracts geometri- erties identification. Spectroscopic non-destructive
cal features, and employs a support vector machine methods show promise in assessing adulteration, fun-
(SVM) for classification. It also grades grains based gal infection, and quality (Natarajan and Ponnusamy,
on milling defects, providing comprehensive quality 2023). Computer vision is crucial in seed testing,
assessment (Mittal et al., 2019). In India, where rice particularly for seed and seedling classification. This
is a staple for 70% of the population, food quality review explores the challenges in seed identification,
is a crucial concern. This paper addresses the issue the limitations of current techniques, and the poten-
of rice quality and proposes using image analysis to tial of deep learning. It recommends optimizing image
ensure accurate rice size and, therefore, protein con- acquisition, dataset construction, and model develop-
tent. Feature extraction techniques and ML models ment for seed identification (Zhao et al., 2022).
are used to assess rice quality (Panigrahi, 2020). A Rice quality assessment benefits from modern high-
study evaluates machine vision techniques for classi- precision instruments and agricultural AI. The detec-
fying six Asian rice varieties. tion of rice appearance quality with high-precision
Digital images captured in open field conditions are instruments has gained traction in agriculture (He et
processed, and various features are extracted. These al., 2023). Nutrient deficiency affects crop production
features are used to discriminate rice varieties, achiev- significantly. Computer vision and ML technologies
ing high classification accuracies with different classi- are employed to detect nutrient deficiencies in crops.
fiers (Qadri et al., 2021). Micronutrient malnutrition This overview explores recent research in crop nutri-
affects billions of people worldwide. Enhancing min- ent content identification and its challenges (Sudhakar,
eral concentrations in crops through biofortifica- M., and R. M. Priya, 2023).
tion is a sustainable solution. Quality assessment of Machine vision plays a vital role in plant pheno-
grains is traditionally manual, time-consuming, and typing, offering non-destructive solutions for trait
variable. This paper proposes using image process- estimation and classification. This comprehensive
ing to analyze grain quality, considering physical and survey outlines various imaging methods along with
chemical characteristics. Edge detection is employed their applications in plant phenotyping, including
to determine grain boundaries (Velavan et al., 2021). deep learning algorithms (Kolhar and Jayant, 2023).
Automated rice variety identification is a challenging Automated detection and classification of rice crop
task, requiring expertise in agriculture and advanced diseases are crucial for improving crop output. This
AI technology. To address this challenge, an automatic system uses computer vision, ML, image process-
rice variety identification system has been developed, ing, and DL to identify diseases like brown leaf spot,
Applied Data Science and Smart Systems 605
ANN-based classification
The biological nervous system, which comprises sev-
eral nodes and replicates biological neurons in the
human brain, served as the model for the ANN clas-
sification system. The nodes, or neurons, are inter-
connected and engage in communication. The nodes
receive the input data, analyze it, and then transform
it before sending it across the connection to other
neurons as an output. Each connection will have a
weight that may be changed. Every node’s output is
processed via the weight. A hidden layer is a layer that
resides between the input and the output and con-
ducts calculations on weighted inputs to generate the
net input.
The real output is subsequently created by combin-
ing the activation function with the net input. Input
and output nodes added together may equal one or
up to two hidden layers. The network randomly adds
weight while passing from the input to the output Figure 78.2 Flow chart of the process
layer, and the result is sent to the next nodes. The ulti-
mate result is compared with the objective. If the out-
put does not match the goal, it propagates backward Table 78.1 Grade of rice grains
and modifies the weights. The architecture of an ANN Grade Type
is shown in Figure 78.1. Figure 78.2 illustrates a sug-
gested automated rice grain quality evaluation system 1 Perfect small rice grains
based on ANN. 2 Perfect large rice grains
Matta and Ponni, two widely consumed rice types,
3 Rice grains with impurities
are the subjects of this research. As seen in Table 78.1,
there are four grades assigned to them. The rice is 4 Imperfect (mixed) rice grains
divided into two categories, Ponni rice, and Matta
rice, using the mean RGB values of various images.
Figure 78.3 for the Ponni variety and Figure 78.4 for grains are divided into grades 1, 2, 3, and 4 based on
the Matta type of rice show the test findings. The rice geometrical characteristics.
606 Artificial intelligence and machine vision-based assessment of rice seed quality
Results
The NN-toolbox of MATLAB is used to carry out
the classification scheme for rice grains. The NN
classifier system has utilized seven data sets. The
ANN framework employed in this study is shown in
Figure 78.5. Seven characteristics have been consid-
ered inputs, while rice quality and variety have been Figure 78.5 ANN framework
Applied Data Science and Smart Systems 607
considered outputs. To reach the best level of accuracy, The different rice seeds are classified on the basis
this model comprises 48 hidden layers. Figure 78.6 of their perimeter assessment. Table 78.2 details the
presents the outcomes of the NN classifier and regres- intended value of pixels depending on the particle
sion plots are discussed as shown in Figure 78.7. analysis for every rice seed in one sample image. Table
78.3, differentiates the rice seeds among small, nor-
mal, large, and broken rice granules using the object
detection method. The table also presents the values
calculated in percentage depending on the total seeds
in each sample using the machine-based system analy-
sis. Further, Table 78.4 states the values of different
rice seeds as detected by a human inspector, with the
corresponding percentage values concerning the total
number of seeds present in the sample.
Tables 78.3–78.5 presents the values of the number
of different quality rice seeds by human detection and
using the machine vision method. From the results, it
can be seen that more accurate results are produced
when a system based on machine vision is used as
compared to manual detection. An error analysis, as
presented in Table 78.5, showed a major variation
in the percentage of error when detected manually
1 198
2 198
3 176
4 199
5 198
6 236
Figure 78.6 Grading of rice grains according to variet-
7 227
ies using ANN classifier
8 114
9 177
10 121
11 206
12 190
13 209
14 209
15 116
16 203
17 223
18 226
19 187
20 225
21 237
22 211
23 148
24 216
25 196
26 218
Figure 78.7 Regression plots
608 Artificial intelligence and machine vision-based assessment of rice seed quality
Table 78.3 Results of a random sample regarding the size of the rice seeds with percentage values
Table 78.4 Results of a random sample in terms of the size of the rice seeds with percentage values by Human inspector
compared to the error using the machine-based sys- Kolhar, S. and Jayant, J. (2023). Plant trait estimation and clas-
tem for normal and chalky seeds. sification studies in plant phenotyping using machine
vision–A review. Inform. Proc. Agricul., 10(1), 114–135.
Komal, Komal, Ganesh Kumar Sethi, and Rajesh Kumar
Conclusion Bawa. (2022). A prototype of automatic rice variety
Today, customers are becoming more and more identification system using artificial intelligence tech-
niques. In AIP Conference Proceedings, 2455(1). AIP
health-conscious and hence are concerned with the
Publishing. [Link]
quality of food they consume. To make sure the rice
Mittal, S., Dutta, M. K., and Issac, A. (2019). Non-destruc-
grains are of good quality, an AI-based system is pre- tive image processing-based system for assessment of
sented here used for assessing the quality of the rice rice quality and defects for classification according to
grains in terms of grades. At the same time, a machine inferred commercial value. Measurement, 148, 106969.
vision-based system was used to differentiate as per Singh, S., Singh, J., Goyal, S. B., Sehra, S. S., Ali, F., Alkhafa-
the rice grain size. The AI system uses an ANN clas- ji, M. A., and Singh, R. (2023). A novel framework to
sifier to differentiate the grade rice using the different avoid traffic congestion and air pollution for sustain-
geometrical and morphological aspects. The clas- able development of smart cities. Sustain. Ener. Tech-
sifier’s overall efficiency was 83%, whereas a 10% nol. Assess., 56, 103125. [Link]
error rate was estimated using the manual system of seta.2023.103125.
Mochida, K., Koda, S., Inoue, K., Hirayama, T., Tanaka, S.,
differentiating different rice grains. This research can
Nishii, R., and Melgani, F. (2019). Computer vision-
also be extended further by using more parameters to
based phenotyping for improvement of plant produc-
increase the machine’s accuracy. Also, another system tivity: a machine learning perspective. Giga Sci., 8(1),
can be made where rice granules are detected simulta- giy153.
neously for both their size and grade quality. Natarajan, S. and Ponnusamy, V. (2022). A review on rice
quality analysis. Soft comput. Sec. Appl. Proc. ICSCS,
References 2022, 119–133.
Panigrahi, J., Pattnaik, P., Dash, B. B., and Dash, S. R.
Al-Shayea, Q. K., and Bahia, I. S. H. (2010). Urinary system (2020). Rice quality prediction using computer vision.
diseases diagnosis using artificial neural networks. Int. Int. Conf. Comp. Sci. Engg. Appl. (ICCSEA), 1–5.
J. Comp. Sci. Netw. Sec. (IJCSNS), 10(7), 118–122. Patel, B. and Aakanksha, S. (2023). Rice variety classifica-
Avudaiappan, T., S. Sangamithra, A. S. Roselin, S. S. Farha- tion & yield prediction using semantic segmentation
na, and K. M. Visalakshi. (2019). Analysing rice seed of agro-morphological characteristics. Multimed. Tool
quality using machine learning algorithms. SSRG In- Appl., 1–18.
ternational Journal of Computer Science and Engineer- Qadri, S., Aslam, T., Nawaz, S. A., Saher, N., Razzaq, A., Ur
ing (SSRG—IJCSE)—Special Issue ICRTCRET 474. Rehman, M., Ahmad, N., Shahzad, F., and Qadri, S. F.
Dhiman, P., Kukreja, V., Manoharan, P., Kaur, A., Kamruz- (2021). Machine vision approach for classification of
zaman, M. M., Dhaou, I. B., and Iwendi, C. (2022). A rice varieties using texture features. Int. J. Food Prop.,
novel deep learning model for detection of severity lev- 24(1), 1615–1630.
el of the disease in citrus fruits. Electronics, 11(3), 495. Stubbs, Megan. (2016). Big data in US agriculture. Wash-
Haridasan, A., Thomas, J., and Raj, E. D. (2023). Deep ington, DC: Congressional Research Service.
learning system for paddy plant disease detection and Sudhakar, M., and R. M. Priya. (2023). Computer Vision
classification. Environ. Monit. Assess., 195(1), 120. Based Machine Learning and Deep Learning Ap-
Hasija, T., Kadyan, V., Guleria, K., Alharbi, A., Alyami, H., proaches for Identification of Nutrient Deficiency
and Goyal, N. (2022). Prosodic feature-based dis- in Crops: A Survey. Nature Environment & Pol-
criminatively trained low resource speech recognition lution Technology. 22(3), 1387–1399. [Link]
system. Sustainability, 14(2), 614. org/10.46488/NEPT.2023.v22i03.025
He, Y., Fan, B., Sun, L., Fan, X., Zhang, J., Li, Y., and Suo, X. Trivedi, N. K., Gautam, V., Anand, A., Aljahdali, H. M., Vil-
(2023). Rapid appearance quality of rice based on ma- lar, S. G., Anand, D., Goyal, N., and Kadry, S. (2021).
chine vision and convolutional neural network research Early detection and classification of tomato leaf dis-
on automatic detection system. Front. Plant Sci., 14. ease using high-performance deep neural network.
Hidayat, S. S., Rahmawati, D., Prabowo, M. C. A., Triyono, Sensors, 21(23), 7987.
L., and Putri, F. T. (2023). Determining the rice seeds Velavan, P., S. Keerthana, and A. Mercy Shanthi Rani.
quality using convolutional neural network. JOIV Int. (2021). Rice Quality Analysis Using Edge Detection
J. Inform. Visual., 7(2), 527–534. Algorithm. Turkish Online Journal of Qualitative In-
Ji, B., Sun, Y., Yang, S., and Wan, J. (2007). Artificial neural quiry, 12(6), p6630.
networks for rice yield prediction in mountainous re- Vignesh, and Elakya. (2022). Identification of unhealthy leaves
gions. J. Agricul. Sci., 145(3), 249–261. in paddy by using computer vision based deep learning
Kadyan, V., Hasija, T., and Singh, A. Prosody features based model. Int. J. Elec. Electron. Res., 10(4), 796–800.
low resource Punjabi children ASR and T-NT classi- Zhao, L., Haque, S. M., and Wang, R. (2022). Automated
fier using data augmentation. (2023). Multimed. Tools seed identification with computer vision: Challenges
Appl., 82(3), 3973–3994. and opportunities. Seed Sci. Technol., 50(2), 75–102.