<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD with MathML3 v1.1d2 20140930//EN" "JATS-journalpublishing1-mathml3.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" dtd-version="1.1d2" xml:lang="en">
  <front>
    <journal-meta>
      <journal-id journal-id-type="nlm-ta">JIAP</journal-id>
      <journal-id journal-id-type="publisher-id">IECE</journal-id>
      <journal-title-group>
        <journal-title>IECE Journal of Image Analysis and Processing</journal-title>
      </journal-title-group>
      <issn pub-type="ppub" publication-format="print">pending</issn>
      <issn pub-type="epub" publication-format="electronic">pending</issn>
      <publisher>
        <publisher-name>Institute of Emerging and Computer Engineering Inc</publisher-name>
        <publisher-loc>522 W RIVERSIDE AVE STE N, SPOKANE, WA, 99201-0508, UNITED STATES</publisher-loc>
      </publisher>
    </journal-meta>
    <article-meta>
      <article-id pub-id-type="doi">10.62762/JIAP.2025.227089</article-id>
      <article-categories>
        <subj-group subj-group-type="heading">
          <subject>Research Article</subject>
        </subj-group>
      </article-categories>
      <title-group>
        <article-title>Plant Disease Detection Using Deep Learning Techniques</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <name>
            <surname>Zainab</surname>
            <given-names>Zunaira</given-names>
          </name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <contrib-id contrib-id-type="orcid">https://orcid.org/0000-0003-1983-8201</contrib-id>
          <name>
            <surname>Mahum</surname>
            <given-names>Rabbia</given-names>
          </name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
      </contrib-group>
      <author-notes>
        <corresp id="cor2">Corresponding Author: Rabbia Mahum. Email: <email>rabbia.mahum@uettaxila.edu.pk</email></corresp>
      </author-notes>
      <pub-date date-type="pub" pub-type="epub" publication-format="online">
        <day>20</day>
        <month>3</month>
        <year>2025</year>
      </pub-date>
      <volume>1</volume>
      <issue>1</issue>
      <fpage>36</fpage>
      <lpage>44</lpage>
      <history>
        <date date-type="received">
          <day>23</day>
          <month>1</month>
          <year>2025</year>
        </date>
        <date date-type="accepted">
          <day>10</day>
          <month>3</month>
          <year>2025</year>
        </date>
      </history>
      <permissions>
        <copyright-statement>© 2025 by the Authors. Published by Institute of Emerging and Computer Engineers. This is an open access article under the CC BY license (https://creativecommons.org/licenses/by/4.0/).</copyright-statement>
        <copyright-year>2025</copyright-year>
        <copyright-holder>Institute of Emerging and Computer Engineering Inc</copyright-holder>
        <license xlink:href="https://creativecommons.org/licenses/by/4.0/">
        <license-p>This work is licensed under a <ext-link ext-link-type="uri" xlink:type="simple" xlink:href="https://creativecommons.org/licenses/by/4.0/">Creative Commons Attribution 4.0 International License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.</license-p>
        </license>
      </permissions>
      <self-uri xlink:href="https://www.iece.org/article/abs/jiap.2025.227089">This article is available from https://www.iece.org/article/abs/jiap.2025.227089</self-uri>
      <abstract>
        <p>Plant diseases create one of the most serious risks to the world’s food supply, reducing agricultural production and endangering millions of people’s lives. These illnesses can destroy crops, disrupt food supply networks, and increase the danger of food deficiency, emphasizing the importance of establishing strong methods to protect the world’s food sources. The approaches of deep learning have transformed the field of plant disease diagnosis, providing sophisticated and perfect solutions for early detection and management. However, a prevalent concern with deep learning models is their susceptibility to a lack of generalization and robustness when faced with novel crop and disease categories that were not included in the training dataset. To tackle this problem, this study present a novel deep learning-based model that can differentiate between healthy and diseased leaves in various crops, even if the model was not trained on them. The main idea is to identify the diseased small leaf regions instead of the diseased leaf’s overall appearance and to calculate the disease’s prevalence rate on the entire leaf. To achieve efficient classification and to take advantage of the Inception model’s superiority in disease recognition, this study use a small Inception model architecture, which can process small regions without sacrificing performance. This study trained and evaluated the proposed approach using the highly regarded PlantVillage dataset, which is the most used dataset because of its extensive and varied coverage, in order to verify its efficacy. The accuracy percentage of the proposed approach was 99.75%. This novel method tackles the crucial problem of model generalization to a variety of crops and diseases in addition to improving the precision of plant disease diagnosis. Furthermore, it demonstrated its potential for wide applicability and support to global food security programs by outperforming the current methodologies in its capacity to identify any illness across any crop type.</p>
      </abstract>
      <kwd-group kwd-group-type="author" xml:lang="en">
        <kwd>deep learning</kwd>
        <kwd>agriculture</kwd>
        <kwd>plant disease</kwd>
        <kwd>disease detection</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="S1">
      <label>1.</label>
      <title>Introduction</title>
      <p id="S1.p1">Agriculture is a vital part of all economic systems and a basic source of food that ensures life will continue. To maximize output and raise quality, the agriculture industry must be improved. This entails creating the ideal environment for crops and plants to develop healthily. Plant degradation is typically caused by diseases. The Food and Agriculture Organization of the United Nations (FAO) estimates that the annual cost of these diseases to the world economy is around $220 billion. They cause harm to crops and, occasionally, their complete destruction. In fact, viruses, bacteria, fungi, and microscopic animals attack plants, altering their original form and affecting their important functions [<xref rid="ref001" ref-type="bibr">1</xref>].</p>
      <p id="S1.p2">Plants can be saved when infections are detected early and neutralized. To protect crop health and ensure ongoing agricultural productivity, early diagnosis of plant diseases is essential. However, manual inspection is a major component of traditional systems, which makes them unreliable, time-consuming, and unsuitable for large-scale farming. Due to a lack of experience, farmers frequently find it difficult to detect infections early on, which causes delays in interventions and large crop losses. Given that agriculture is the foundation of many economies and food security systems, automated, precise, and effective disease detection methods are desperately needed. Deep learning models have emerged as a promising solution to these problems.</p>
      <p id="S1.p3">The crops are safeguarded, and losses are prevented to a greater extent the sooner they are identified. Due to a lack of expertise and experience, the conventional methods of disease detection—which are primarily reliant on human diagnosis—are insufficient and time-consuming [<xref rid="ref002" ref-type="bibr">2</xref>]. The diagnosis must be based on a more trustworthy technique because the data collection method and verification frequency are also insufficient. Modern technologies have been provided as an automatic way to identify plant diseases for this purpose [<xref rid="ref003" ref-type="bibr">3</xref>]. Among these technologies, advanced instruments like sensors, drones, and robots have transformed farmers’ field management [<xref rid="ref004" ref-type="bibr">4</xref>], and machine learning has opened up new avenues for data analysis [<xref rid="ref005" ref-type="bibr">5</xref>].</p>
      <p id="S1.p4">With encouraging outcomes, the use of deep learning and machine learning approaches in plant disease identification is a quickly developing subject [<xref rid="ref006" ref-type="bibr">6</xref>]. However, because deep learning-based techniques rely on automatic feature extraction rather than human feature selection, they have demonstrated their efficacy above other machine learning techniques, particularly in relation to image identification [<xref rid="ref007" ref-type="bibr">7</xref>]. DL-based methods for the Identification and detection of plant diseases have been proposed in a number of research studies. However, there are a lot of barriers that keep this technology from being used more effectively. The impossibility of gathering dataset photos for every illness across all leaf kinds is, in fact, one barrier. Furthermore, a number of illnesses are spreading quickly, making it difficult to catch them on leave on time.</p>
      <p id="S1.p5">Furthermore, the study [<xref rid="ref008" ref-type="bibr">8</xref>] demonstrated that the system is impacted by both the crop’s and the disease’s properties during the learning process. This implies that features taken from one crop and illness cannot be transferred to other crops and diseases. Classification model training on a dataset comprising pairs of various crops and diseases is necessary to create a generalized model for classifying plant diseases. Unfortunately, there isn’t a dataset like that, and creating one is difficult, if not impossible. On the planet, there are millions of crops and plants, and millions of diseases could harm them.</p>
      <p id="S1.p6">Even though deep learning has demonstrated remarkable outcomes in the Identification of plant diseases, current models frequently lack generalization and resilience. The majority of methods are trained on certain datasets and have trouble detecting illnesses in new crops or in different environmental settings. This constraint occurs because small, disease-affected areas are not given as much attention to conventional models as the overall appearance of the leaf. Developing dependable, scalable solutions for actual agricultural settings is made more difficult by the lack of labeled datasets that reflect a variety of crop kinds and disease categories. To overcome these constraints, a model that can accurately identify diseases in a variety of crops and environments is needed.</p>
      <p id="S1.p7">This study presents a unique deep learning-based approach that aims to address this obstacle by generalizing the method for identifying plant diseases in various kinds of plants, hence overcoming the limits of specific crop training difficulties reported in the existing research. Our method is centered on the primary goal of diagnosing the disease, not just the sick leaf’s outward look. It places special emphasis on identifying the healthiness or illness in tiny leaf fragments. This study produced a method to achieve this, which involves dividing each image of leaves in the training dataset into smaller patches, so separating the disease from the healthy parts. By using this dividing approach, it is easier to extract disease-related information from the crops and enhance healthy data.</p>
    </sec>
    <sec id="S2">
      <label>2.</label>
      <title>Related Work</title>
      <p id="S2.p1">Afsharpour et al. [<xref rid="ref009" ref-type="bibr">9</xref>] and Albahar [<xref rid="ref010" ref-type="bibr">10</xref>] claim that DL models enhance fruit disease diagnosis and boost agricultural productivity. For the detection of fruit rot, the approach suggested in [<xref rid="ref010" ref-type="bibr">10</xref>] was more reliable. By altering factors like lighting and picture obstructions, the suggested approach was able to overcome the consequences.</p>
      <p id="S2.p2">Huang et al. [<xref rid="ref011" ref-type="bibr">11</xref>] proposed neural network enhanced multi-scale feature extraction and citrus fruit disease detection by combining the form of the inception module with the cutting-edge EfficientNetV2. The VGG was used instead of the U-Net backbone to enhance the network’s segmentation performance. The proposed neural network, which was developed by combining the Inception module with the cutting-edge EfficientNetV2, enhanced multi-scale feature extraction and citrus fruit disease detection. Lokesh et al. [<xref rid="ref012" ref-type="bibr">12</xref>] proposed an artificial intelligence strategy for identifying leaf sickness that incorporates generative adversarial networks (GAN) and deep learning. The strategy surpassed current techniques, achieving higher classification accuracy for healthy and diseased leaf images. Wang et al. [<xref rid="ref013" ref-type="bibr">13</xref>] improved the design of ResNet by suggesting a multi-scale ResNet. The system is used to identify vegetable leaf diseases on constrained hardware by lowering the model’s size. The multi-scale ResNet has outperformed other systems (VGG16, AlexNet, SqueezeNet, and ResNet50). Ren et al. [<xref rid="ref014" ref-type="bibr">14</xref>] proposed a new system, the deconvolution-guided VGG network, for detecting tomato leaf illnesses and segmenting disease areas. The proposed model is capable of dealing with all image difficulties, including low light, shadows, and occlusion. Guo et al. [<xref rid="ref015" ref-type="bibr">15</xref>] created a customized tomato leaf disease identification system for Android mobile devices. This method easily differentiates between different diseases at various stages by altering the steps of the AlexNet design, resulting in the Multi-Scale AlexNet.</p>
      <p id="S2.p3">Using an improved version of MobileNetV3, Qiaoli et al. [<xref rid="ref016" ref-type="bibr">16</xref>] described a real-time, non-destructive method for detecting tomato illnesses. The method addresses the problem of model overfitting brought on by insufficient data by optimizing MobileNetV3. In order to differentiate between healthy and diseased leaves, Khalid et al. [<xref rid="ref017" ref-type="bibr">17</xref>] concentrated on training a DL model, namely the YOLOv5 system. Both public and exclusive datasets were then used to test the trained system. The results show how quickly and precisely the model can detect even tiny disease areas on plant leaves. Panchal et al. [<xref rid="ref018" ref-type="bibr">18</xref>] suggested a strategy for detecting crop diseases that includes segmenting images, labeling diseased leaves, pixel-based operations, feature extraction, and convolutional neural networks (CNN) for disease classification. To identify and segment tomato plant illnesses, Shoaib et al. [<xref rid="ref019" ref-type="bibr">19</xref>] combined semantic segmentation methods (U-Net and Modified U-Net) with deep learning (Inception Net).</p>
    </sec>
    <sec id="S3">
      <label>3.</label>
      <title>Methodology</title>
      <p id="S3.p1">The suggested model, shown in Figure <xref ref-type="fig" rid="F1">1</xref>, creates four classes for assessing and estimating the suggested frame by utilizing four well-known transfer learning techniques: ResNet152, VGG19, DenseNet169, and MobileNetv3. After being subjected to four different transfer learning techniques, the data is analyzed and split into two sets: 80% for training and 20% for testing. This division is essential for generalizability assessment, model performance validation, and training. The suggested model shows reliability in a range of situations. In this work, we extend the dataset to train a deep learning model in plant disease diagnosis by the use of image augmentation, a crucial method utilizing Keras’ Image Data Generator. The model is exposed to a greater variety of variations by generating altered copies of images with rotations, zooming, and flipping, which enhances the model’s capacity to handle new data. This is essential to simulate the variability found in medical imaging and strengthen the model’s resistance to noise and fluctuations. The ultimate objective is to create a robust and dependable deep learning model, particularly in the medical field where data is scarce and case diversity and unforeseen conditions are critical. During model training, changes like zooms, flips, shifts, and rotations are added to help create balanced classes.</p>
      <p>
        <fig id="F1">
          <label>Figure 1.</label>
          <caption>
            <p>Proposed Model.</p>
          </caption>
          <graphic xlink:href="fig1.jpg"/>
        </fig>
      </p>
      <p id="S3.p2">This augmentation technique helps create a training dataset that is more varied and extensive, which improves the model’s ability to generalize in a wide range of situations. There are two benefits to using ImageDataGenerator while training a model. In order to strengthen the deep learning model’s ability to recognize complex patterns and features, it first makes sure that the model is exposed to a wider range of training data. Furthermore, the automatic production of enhanced images improves the resilience of the model by reducing its susceptibility to overfitting and increasing its ability to adjust to a wide range of input variables. This augmentation-driven method has been shown to be effective in raising deep learning models’ overall performance, which improves accuracy and durability in practical applications [<xref rid="ref020" ref-type="bibr">20</xref>].</p>
      <sec id="S3.SS1">
        <label>3.1</label>
        <title>Transfer learning model evaluation</title>
        <p id="S3.SS1.p1">A machine learning technique called transfer learning allows a model to be learned for one job and then applied to another that is related to it. This approach shortens the effort and time needed to create high-performing systems, particularly for challenging jobs like natural language processing and picture recognition. Through fine-tuning an established model’s weights, scientists can effectively address novel problems. By utilizing the knowledge gathered from a sizable dataset during the first training phase, transfer learning enables the system to efficiently handle novel difficulties.</p>
        <p id="S3.SS1.p2">This method differs with the usual way of training a system from the ground up, which can be resource intensive and time-consuming. In several fields, such as speech recognition, picture identification and natural language processing, transfer learning has shown promise, particularly in situations with little training.</p>
      </sec>
      <sec id="S3.SS2">
        <label>3.2</label>
        <title>ResNet152</title>
        <p>
          <fig id="F2">
            <label>Figure 2.</label>
            <caption>
              <p>Architecture of RESNET-152.</p>
            </caption>
            <graphic xlink:href="fig2.jpg"/>
          </fig>
        </p>
        <p id="S3.SS2.p1">Microsoft Research created ResNet-152, DCNN architecture with 152 layers. The main advancement in it is the use of residual connections or skip connections, which allow the network to learn residual functions and ease the training of exceptionally deep networks [<xref rid="ref021" ref-type="bibr">21</xref>]. ResNet-152 is effective at tasks like object recognition and image classification because of its depth, which enables it to excerpt complex patterns and features from data. When combined with skip connections, this architectural depth solves the vanishing gradient issue and makes it simple to train incredibly deep networks, as seen in Figure <xref ref-type="fig" rid="F2">2</xref>.</p>
      </sec>
      <sec id="S3.SS3">
        <label>3.3</label>
        <title>Visual geometry group 19</title>
        <p id="S3.SS3.p1">An advancement over the original VGG16 design is the DCNN (deep convolutional neural network) method called VGG19. There are sixteen convolutional layers and three completely linked layers, totaling 19 layers in this model. VGG19’s deep structure, which employs 3×3 convolutional filters for feature extraction, allows it to extract complex patterns and features from picture data [<xref rid="ref022" ref-type="bibr">22</xref>]. By reducing the spatial dimensions of input, max-pooling layers minimize computing complexity. Due to the entire connectivity of the final layers, predictions derived from high-level characteristics taken out by the convolutional layers are possible. For nonlinearity, VGG19 employs the ReLU (Rectified Linear Unit) activation function. Often employed in image classification. In terms of efficiency and performance, it has been outperformed by more recent designs, such as ResNet and Inception, despite their simplicity and depth. Figure <xref ref-type="fig" rid="F3">3</xref> shows the VGG19 architecture.</p>
        <p>
          <fig id="F3">
            <label>Figure 3.</label>
            <caption>
              <p>Architecture VGG19.</p>
            </caption>
            <graphic xlink:href="fig3.jpg"/>
          </fig>
        </p>
      </sec>
      <sec id="S3.SS4">
        <label>3.4</label>
        <title>DenseNet169</title>
        <p id="S3.SS4.p1">A CNN architecture called DenseNet169 was created to address issues with gradient flow and feature reuse in deep networks. In order to restrict the feature’s growth map and minimize its spatial dimensions, transition layers are used in between dense blocks. The Global average pooling layer, which lowers the number of parameters and improves expansions, is frequently used in DenseNet designs. DenseNet169 has proven to perform well in challenges involving image classification. The architecture of DenseNet169 is shown in Figure <xref ref-type="fig" rid="F4">4</xref>.</p>
        <p>
          <fig id="F4">
            <label>Figure 4.</label>
            <caption>
              <p>DenseNet169.</p>
            </caption>
            <graphic xlink:href="fig4.jpg"/>
          </fig>
        </p>
      </sec>
      <sec id="S3.SS5">
        <label>3.5</label>
        <title>MobileNetv3</title>
        <p id="S3.SS5.p1">A neural network architecture called MobileNetV3 was created for edge and mobile devices with limited processing capacity. This third version of the MobileNet series prioritizes accuracy, efficiency, and speed. Two variations, MobileNetV3-Large and MobileNetV3-Small, lightweight inverted residuals and resource-efficient construction blocks are among the key features [<xref rid="ref023" ref-type="bibr">23</xref>]. The architecture of MobileNetv3 is shown in Figure <xref ref-type="fig" rid="F5">5</xref>.</p>
        <p>
          <fig id="F5">
            <label>Figure 5.</label>
            <caption>
              <p>Architecture of MobileNetv3.</p>
            </caption>
            <graphic xlink:href="fig5.jpg"/>
          </fig>
        </p>
      </sec>
    </sec>
    <sec id="S4">
      <label>4.</label>
      <title>Experiments</title>
      <p id="S4.p1">A particular dataset was used to train the Transfer Learning (TL) model, and a matching test dataset was used to assess it. Keras, Sklearn, and TensorFlow cooperate to ensure the model’s success. For high-end systems, 128 block size was ideal. Both the test and train sets were subjected to cross-entropy loss. The results are reported in Table <xref rid="T1" ref-type="table">1</xref>. The DenseNet169 model behaved differently than the VGG19, MobileNetv3, and ResNet152 models, which all had close training and validation losses.</p>
      <p id="S4.p2">This study highlights how essential it is to use a Deep Learning model for the Identification of plant disease. 20% of each class was chosen at random from 14,059 photos in the PlantVillage dataset that were tested. The model classified leaves as either healthy or ill with an accuracy rate of above 95%. However, some photos were incorrectly categorized because of dirt or poor illumination.</p>
      <p id="S4.p3">The Adam optimizer was utilized in the study to train 50 models using cross-entropy loss. The results indicated that whereas DenseNet169 had a different pattern, with training loss reduced and validation loss increasing, VGG19, MobileNetv3, and ResNet152, models had near validation and training losses.</p>
      <p id="S4.p4">MobileNetv3 had the best testing accuracy of 99.75% among the four transfer learning models that were assessed. DenseNet169, on the other hand, displayed oscillations, suggesting sensitivity to changes in the dataset. The accuracy of ResNet152 was continuously high at 98.86%, but MobileNetv3 outperformed it due to its lightweight design and effective feature extraction. These findings demonstrate how crucial model architecture is for striking a balance between generalization, accuracy, and computational economy. VGG19 and ResNet152 models had training and validation losses, but MobileNetv3 demonstrated a steady training process with few oscillations. ResNet152 demonstrated the best performance among the four transfer learning models assessed in the study, with a validation loss of 0.0241 and 98.86%.</p>
      <p>
        <table-wrap id="T1">
          <label>Table 1</label>
          <caption>
            <p>Results from four different transfer learning models More than 50 Epochs.</p>
          </caption>
          <table>
            <thead>
              <tr>
                <th style="border-top: 1px solid black;" rowspan="2" align="center">
                  <p>
                    <table-wrap>
                      <table>
                        <tr>
                          <td align="center">
                            <bold>Architecture</bold>
                          </td>
                        </tr>
                      </table>
                    </table-wrap>
                  </p>
                </th>
                <th style="border-top: 1px solid black;" colspan="2" align="center">
                  <bold>Testing</bold>
                </th>
                <th style="border-top: 1px solid black;" colspan="2" align="center">
                  <bold>Training</bold>
                </th>
              </tr>
              <tr>
                <th style="border-top: 1px solid black;" align="center">
                  <bold>Loss</bold>
                </th>
                <th style="border-top: 1px solid black;" align="center">
                  <bold>Acc</bold>
                </th>
                <th style="border-top: 1px solid black;" align="center">
                  <bold>Loss</bold>
                </th>
                <th style="border-top: 1px solid black;" align="center">
                  <bold>Acc</bold>
                </th>
              </tr>
            </thead>
            <tbody>
              <tr>
                <td style="border-top: 1px solid black;" align="center">VGG19</td>
                <td style="border-top: 1px solid black;" align="center">0.124</td>
                <td style="border-top: 1px solid black;" align="center">95.62</td>
                <td style="border-top: 1px solid black;" align="center">0.045</td>
                <td style="border-top: 1px solid black;" align="center">99.07</td>
              </tr>
              <tr>
                <td align="center">ResNet152</td>
                <td align="center">0.185</td>
                <td align="center">96.92</td>
                <td align="center">0.0603</td>
                <td align="center">98.86</td>
              </tr>
              <tr>
                <td align="center">MobileNetv3</td>
                <td align="center">0.127</td>
                <td align="center">98.52</td>
                <td align="center">0.0359</td>
                <td align="center">99.75</td>
              </tr>
              <tr>
                <td style="border-bottom: 1px solid black;" align="center">DenseNet169</td>
                <td style="border-bottom: 1px solid black;" align="center">0.958</td>
                <td style="border-bottom: 1px solid black;" align="center">97.53</td>
                <td style="border-bottom: 1px solid black;" align="center">0.0241</td>
                <td style="border-bottom: 1px solid black;" align="center">99.22</td>
              </tr>
            </tbody>
          </table>
        </table-wrap>
      </p>
      <p id="S4.p5">DenseNet169’s training and validation accuracy fluctuated the most, despite having the lowest accuracy. Promising outcomes were also displayed by VGG19 and MobileNetv3, however there was some fluctuation in validation accuracy.</p>
      <sec id="S4.SS1">
        <label>4.1</label>
        <title>Results and discussion of the experiment</title>
        <p id="S4.SS1.p1">The suggested transfer learning systems use the confusion matrix to evaluate their results, which includes metrics such as recall, precision, accuracy, and F1 score. Using common criteria, such as accuracy, precision, recall, and F1 score, the model’s performance was assessed. Precision quantifies the proportion of identified diseased leaves that were correctly diagnosed, whereas accuracy indicates the overall correctness of forecasts. The F1 score offers a compromise between precision and recall, whereas recall shows how successfully the model detected every diseased leaf. The confusion matrix, which is commonly a square matrix, provides a thorough perspective of model outcomes. Table <xref rid="T2" ref-type="table">2</xref> shows the confusion matrix, with FP denoting false positives, FN denoting false negatives, and TP denoting true positives. The F1 score is calculated as the harmonic means of recall and precision.</p>
        <p id="S4.SS1.p2">This study evaluated the proposed model’s performance using multiple indicators and provided the findings. The model’s resilience is demonstrated by the low false positive and false negative rates, while the high true positive rates in every category show that it successfully detects a variety of disease types. The model’s capacity for generalization is further supported by its excellent performance across several classes. A useful tool for evaluating the model’s classification results is the confusion matrix. This matrix categorizes diseases from 1 to 3, with each integer identifier representing a specific disease type: 1 for "fungal diseases," 2 for "bacterial diseases," and 3 for "viral diseases." This systematic numbering clearly represents the model’s classification findings. The confusion matrix shows that the MobileNet model outperformed expectations.</p>
        <p id="S4.SS1.p3">These numerical values in the matrix offer important information about how well the model classifies photos of various conditions. The maximum accuracy was attained by ResNet152, which was followed by DenseNet169, MobileNetV3, and VGG19, which all showed respectable accuracy but fell short. These tests, which were carried out after each model was trained for 50 epochs, indicate that ResNet152 performed better than the other systems on a consistent basis. The four models’ different topologies highlight the complex trade-offs between system computational efficiency and accuracy. This study gives important insights into well-informed decision-making by helping to identify their strengths and limitations.</p>
      </sec>
      <sec id="S4.SS2">
        <label>4.2</label>
        <title>Discussion</title>
        <p id="S4.SS2.p1">A thorough summary of the accuracy measurements for each model at both the highest and least epoch counts is given in Table <xref rid="T2" ref-type="table">2</xref>. With the highest validation accuracy of 98.52% at epoch 45 and epoch 50, this study get the greatest training accuracy of 99.75%, and the MobileNet model performs better than any other model. The names Mi_Acc and Mx_Acc stand for Minimum Accuracy and Maximum Accuracy, respectively, and Mi_Ep and Mx_Ep for Minimum Epochs and Maximum Epochs, respectively, to aid in clear communication and notation.</p>
        <p>
          <table-wrap id="T2">
            <label>Table 2</label>
            <caption>
              <p>Accuracy Analysis across Epochs for Various Transfer Learning Methods.</p>
            </caption>
            <table>
              <thead>
                <tr>
                  <th style="border-top: 1px solid black;" align="center">
                    <p>
                      <table-wrap>
                        <table>
                          <tr>
                            <td align="center">
                              <bold>Transfer</bold>
                            </td>
                          </tr>
                          <tr>
                            <td align="center">
                              <bold>learning model</bold>
                            </td>
                          </tr>
                        </table>
                      </table-wrap>
                    </p>
                  </th>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>Stage</bold>
                  </th>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>M×-Ep</bold>
                  </th>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>M×-Acc</bold>
                  </th>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>Mi-Ep</bold>
                  </th>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>Mi-Acc</bold>
                  </th>
                </tr>
              </thead>
              <tbody>
                <tr>
                  <td style="border-top: 1px solid black;" rowspan="2" align="center">VGG19</td>
                  <td style="border-top: 1px solid black;" align="center">Training</td>
                  <td style="border-top: 1px solid black;" align="center">50</td>
                  <td style="border-top: 1px solid black;" align="center">98.78</td>
                  <td style="border-top: 1px solid black;" align="center">1</td>
                  <td style="border-top: 1px solid black;" align="center">70.22</td>
                </tr>
                <tr>
                  <td align="center">Testing</td>
                  <td align="center">50</td>
                  <td align="center">96.91</td>
                  <td align="center">1</td>
                  <td align="center">81.45</td>
                </tr>
                <tr>
                  <td rowspan="2" align="center">DenseNet169</td>
                  <td align="center">Training</td>
                  <td align="center">50</td>
                  <td align="center">98.12</td>
                  <td align="center">1</td>
                  <td align="center">69.34</td>
                </tr>
                <tr>
                  <td align="center">Testing</td>
                  <td align="center">45</td>
                  <td align="center">97.78</td>
                  <td align="center">6</td>
                  <td align="center">80.96</td>
                </tr>
                <tr>
                  <td rowspan="2" align="center">ResNet152</td>
                  <td align="center">Training</td>
                  <td align="center">38</td>
                  <td align="center">99.08</td>
                  <td align="center">1</td>
                  <td align="center">95.42</td>
                </tr>
                <tr>
                  <td align="center">Testing</td>
                  <td align="center">47</td>
                  <td align="center">98.68</td>
                  <td align="center">2</td>
                  <td align="center">90.77</td>
                </tr>
                <tr>
                  <td style="border-bottom: 1px solid black;" rowspan="2" align="center">MobileNetv3</td>
                  <td align="center">Training</td>
                  <td align="center">30</td>
                  <td align="center">99.75</td>
                  <td align="center">1</td>
                  <td align="center">77.12</td>
                </tr>
                <tr>
                  <td style="border-bottom: 1px solid black;" align="center">Testing</td>
                  <td style="border-bottom: 1px solid black;" align="center">46</td>
                  <td style="border-bottom: 1px solid black;" align="center">99.52</td>
                  <td style="border-bottom: 1px solid black;" align="center">1</td>
                  <td style="border-bottom: 1px solid black;" align="center">90.27</td>
                </tr>
              </tbody>
            </table>
          </table-wrap>
        </p>
        <p>
          <table-wrap id="T3">
            <label>Table 3</label>
            <caption>
              <p>Comparison of the suggested and other models’ classification accuracy.</p>
            </caption>
            <table>
              <thead>
                <tr>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>Author</bold>
                  </th>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>Year</bold>
                  </th>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>Dataset</bold>
                  </th>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>Method</bold>
                  </th>
                  <th style="border-top: 1px solid black;" align="center">
                    <bold>Accuracy</bold>
                  </th>
                </tr>
              </thead>
              <tbody>
                <tr>
                  <th style="border-top: 1px solid black;" align="left">Wang et al. [<xref rid="ref013" ref-type="bibr">13</xref>]</th>
                  <th style="border-top: 1px solid black;" align="left">2020</th>
                  <td style="border-top: 1px solid black;" align="left">AI Challenge2018, Self, PlantVillage,</td>
                  <td style="border-top: 1px solid black;" align="left">CNN (Multi-scale ResNet)</td>
                  <td style="border-top: 1px solid black;" align="left">95.95%</td>
                </tr>
                <tr>
                  <th align="left">Ren et al. [<xref rid="ref014" ref-type="bibr">14</xref>]</th>
                  <th align="left">2020</th>
                  <td align="left">PlantVillage</td>
                  <td align="left">CNN (Deconvolution Guided VGGNet)</td>
                  <td align="left">93.25%</td>
                </tr>
                <tr>
                  <th align="left">Guo et al. [<xref rid="ref015" ref-type="bibr">15</xref>]</th>
                  <th align="left">2019</th>
                  <td align="left">PlantVillage, Self</td>
                  <td align="left">CNN (Improved Multi-Scale AlexNet)</td>
                  <td align="left">92.7%</td>
                </tr>
                <tr>
                  <th align="left">Qiaoli et al. [<xref rid="ref016" ref-type="bibr">16</xref>]</th>
                  <th align="left">2022</th>
                  <td align="left">PlantVillage</td>
                  <td align="left">CNN (MobileNetV3)</td>
                  <td align="left">98.25%</td>
                </tr>
                <tr>
                  <th align="left">Khalid et al. [<xref rid="ref017" ref-type="bibr">17</xref>]</th>
                  <th align="left">2023</th>
                  <td align="left">PlantVillage, self</td>
                  <td align="left">CNN (YOLOv5)</td>
                  <td align="left">93%</td>
                </tr>
                <tr>
                  <th align="left">Panchal et al. [<xref rid="ref018" ref-type="bibr">18</xref>]</th>
                  <th align="left">2023</th>
                  <td align="left">PlantVillage</td>
                  <td align="left">CNN (naïve network)</td>
                  <td align="left">90.40%</td>
                </tr>
                <tr>
                  <th align="left">Shoaib et al. [<xref rid="ref019" ref-type="bibr">19</xref>]</th>
                  <th align="left">2022</th>
                  <td align="left">PlantVillage</td>
                  <td align="left">CNN (Inception Net)</td>
                  <td align="left">99.12%</td>
                </tr>
                <tr>
                  <th style="border-bottom: 1px solid black;" align="left">Our proposed model</th>
                  <th style="border-bottom: 1px solid black;" align="left">2024</th>
                  <td style="border-bottom: 1px solid black;" align="left">Plant village</td>
                  <td style="border-bottom: 1px solid black;" align="left">Transfer Learning Approach</td>
                  <td style="border-bottom: 1px solid black;" align="left">99.75%</td>
                </tr>
              </tbody>
            </table>
          </table-wrap>
        </p>
        <p id="S4.SS2.p2">By juxtaposing the current state of the field with the anticipated outcomes of this research, Table <xref rid="T2" ref-type="table">2</xref> and Figure <xref ref-type="fig" rid="F6">6</xref> comprehensively illustrate the evolutionary trajectory and projected advancements in plant disease detection and classification methodologies.</p>
        <p id="S4.SS2.p3">The potential for practical use in agricultural monitoring systems is indicated by the MobileNetv3 outstanding accuracy and stability. Farmers can use the model to detect diseases in real time scenarios using drones or mobile devices, allowing for prompt action, and minimizing crop loss. This useful tool can help farmers better control crop health, reducing yield loss and guaranteeing sustainable farming methods.</p>
        <p>
          <fig id="F6">
            <label>Figure 6.</label>
            <caption>
              <p>Timeline for training four models of transfer learning.</p>
            </caption>
            <graphic xlink:href="fig6.jpg"/>
          </fig>
        </p>
        <p id="S4.SS2.p4">A crucial component of the proposed strategy is the MobileNetv3 model, which has an amazing 99.75% accuracy rate. This demonstrates that the model is a useful tool for disease diagnosis since it can effectively identify and predict the presence of any illness in plant pictures.</p>
        <p id="S4.SS2.p5">The comparison of classification accuracy between the suggested and other current models is shown in Table <xref rid="T3" ref-type="table">3</xref>. The comparison of classification accuracy between the suggested and other cutting-edge techniques is shown in Figure <xref ref-type="fig" rid="F6">6</xref>. With an accuracy of 99.75%, Figure <xref ref-type="fig" rid="F7">7</xref> makes it abundantly evident that the suggested system carried out better than other models.</p>
        <p>
          <fig id="F7">
            <label>Figure 7.</label>
            <caption>
              <p>Comparison of the accuracy of suggested and state-of-the-art methods.</p>
            </caption>
            <graphic xlink:href="fig7.jpg"/>
          </fig>
        </p>
        <p id="S4.SS2.p6">The comparison analysis shows that the suggested MobileNetv3-based model performs better than the current state-of-the-art methods for detecting plant diseases. Table <xref rid="T3" ref-type="table">3</xref> illustrates that the accuracy of the suggested method was higher than that of earlier models, including Multi-scale ResNet (95.95%), Deconvolution-Guided VGGNet (93.25%), and Improved Multi-scale AlexNet (92.7%). The MobileNetv3 model, on the other hand, outperformed all other techniques with an exceptional classification accuracy of 99.75%.</p>
        <p id="S4.SS2.p7">The model’s capacity to generalize across many kinds of crop and disease circumstances while preserving computational efficiency is what makes this work significant. The suggested approach concentrates on identifying discrete diseased patches rather than the overall appearance of the leaf, in contrast to standard models, which frequently struggle with changing environmental circumstances and emerging disease categories. This method improves the model’s resilience and qualifies it for use in practical agricultural applications, such as real-time field monitoring via drones and mobile devices.</p>
        <p id="S4.SS2.p8">Therefore, the suggested method is a useful tool for precision agriculture since it not only increases detection accuracy but also tackles critical issues with model generalization, computational effectiveness, and realistic deployment.</p>
      </sec>
    </sec>
    <sec id="S5">
      <label>5.</label>
      <title>Conclusion</title>
      <p id="S5.p1">The work rigorously evaluates the effectiveness of four different transfer learning systems— VGG19, ResNet152, MobileNetv3, and DenseNet169, on the Plant Village Dataset. The evaluation includes critical performance indicators like precision, accuracy, recall, and f1-score. ResNet152 outperformed all other models in the study with an impressive accuracy of 98.5%. Furthermore, MobileNetv3 has outstanding effectiveness, with 99.75% accuracy, demonstrating its excellent results in plant disease detection. Integrating multiple imaging modalities enhances the proposed model’s performance, durability, and real-world application. Even though the suggested model performed remarkably well on the Plant Village Dataset, still there exist a number of validity risks. The dataset’s controlled conditions may have contributed to the high accuracy attained, and the model’s performance may differ in real-world settings with varying lighting, background noise, and leaf occlusions. Furthermore, the research’s ability to be applied to different agricultural contexts was limited due to its reliance on a single dataset. Early signs that manifest on other plant components, like stems and fruits, may also be missed due to the emphasis on leaf-based disease diagnosis.</p>
      <p id="S5.p2">Future research will concentrate on adding real-world photos from various agricultural settings to the dataset, combining RGB images with multimodal imaging (thermal, hyperspectral, and infrared), and refining the model for deployment on edge devices like smartphones and drones for real-time monitoring in order to overcome these limitations. Future study could look into expanding the suggested model’s usage improving its flexibility. This extension has the ability to expand the influence of models. To summarize, the proposed approach, notably MobileNetv3 and ResNet152, demonstrate significant potential for improving image classification. Further research into various imaging modalities can provide vital insights and improve illness detection applications significantly. Future study will evaluate the efficacy of these models in actual situations using various datasets. The primary objective of this study is to refine image enhancement techniques specifically designed for particular disease groups, thereby facilitating the development of an optimized computational model and a comprehensive, balanced dataset for enhanced diagnostic applications.</p>
    </sec>
  </body>
  <back>
    <ack>
      <title>Acknowledgments</title>
      <p id="ack.p1">This work was supported without any funding.</p>
    </ack>
    <sec id="sec0100" sec-type="COI-statement">
      <title>Conflict of interest</title>
      <p>The authors declare no conflicts of interest. </p>
    </sec>
    <ref-list>
      <title>References</title>
      <ref id="ref001">
        <label>[1]</label>
        <mixed-citation> Dutta, K., Talukdar, D., &amp; Bora, S. S. (2022). Segmentation of unhealthy leaves in cruciferous crops for early disease detection using vegetative indices and Otsu thresholding of aerial images. <italic>Measurement</italic>, 189, 110478. [<uri>https://doi.org/10.1016/j.measurement.2021.110478</uri>] </mixed-citation>
      </ref>
      <ref id="ref002">
        <label>[2]</label>
        <mixed-citation> Li, L., Zhang, S., &amp; Wang, B. (2021). Plant disease detection and classification by deep learning—a review. <italic>IEEE Access</italic>, 9, 56683-56698. [<uri>https://doi.org/10.1109/ACCESS.2021.3069646</uri>] </mixed-citation>
      </ref>
      <ref id="ref003">
        <label>[3]</label>
        <mixed-citation> Ramcharan, A., Baranowski, K., McCloskey, P., Ahmed, B., Legg, J., &amp; Hughes, D. P. (2017). Deep learning for image-based cassava disease detection. <italic>Frontiers in Plant Science</italic>, 8, 1852. [<uri>https://doi.org/10.3389/fpls.2017.01852</uri>] </mixed-citation>
      </ref>
      <ref id="ref004">
        <label>[4]</label>
        <mixed-citation> Attri, I., Awasthi, L. K., &amp; Sharma, T. P. (2023). A review of deep learning techniques used in agriculture. <italic>Ecological Informatics</italic>, 102217. [<uri>https://doi.org/10.1016/j.ecoinf.2023.102217</uri>] </mixed-citation>
      </ref>
      <ref id="ref005">
        <label>[5]</label>
        <mixed-citation> Attri, I., Awasthi, L. K., &amp; Sharma, T. P. (2024). Machine learning in agriculture: A review of crop management applications. <italic>Multimedia Tools and Applications</italic>, 83(5), 12875-12915. [<uri>https://doi.org/10.1007/s11042-023-16105-2</uri>] </mixed-citation>
      </ref>
      <ref id="ref006">
        <label>[6]</label>
        <mixed-citation> Shoaib, M., Shah, B., Khan, S., &amp; Ali, A. (2023). An advanced deep learning models-based plant disease detection: A review of recent research. <italic>Frontiers in Plant Science</italic>, 14, 1158933. [<uri>https://doi.org/10.3389/fpls.2023.1158933</uri>] </mixed-citation>
      </ref>
      <ref id="ref007">
        <label>[7]</label>
        <mixed-citation> Liu, B., Zhang, Y., He, D., &amp; Li, Y. (2017). Identification of apple leaf diseases based on deep convolutional neural networks. <italic>Symmetry</italic>, 10(1), 11. [<uri>https://doi.org/10.3390/sym10010011</uri>] </mixed-citation>
      </ref>
      <ref id="ref008">
        <label>[8]</label>
        <mixed-citation> Yosuke, T. (2019). How Convolutional Neural Networks Diagnose Plant Disease. <italic>Plant Phenomics</italic>. [<uri>https://doi.org/10.34133/2019/9237136</uri>] </mixed-citation>
      </ref>
      <ref id="ref009">
        <label>[9]</label>
        <mixed-citation> Afsharpour, P., Ahmadi, M., &amp; Khosravi, M. (2024). Robust deep learning method for fruit decay detection and plant identification: Enhancing food security and quality control. <italic>Frontiers in Plant Science</italic>, 15, 1366395. [<uri>https://doi.org/10.3389/fpls.2024.1366395</uri>] </mixed-citation>
      </ref>
      <ref id="ref010">
        <label>[10]</label>
        <mixed-citation> Albahar, M. (2023). A survey on deep learning and its impact on agriculture: Challenges and opportunities. <italic>Agriculture</italic>, 13(3), 540. [<uri>https://doi.org/10.3390/agriculture13030540</uri>] </mixed-citation>
      </ref>
      <ref id="ref011">
        <label>[11]</label>
        <mixed-citation> Huang, Z., Li, X., &amp; Wang, Y. (2023). An efficient convolutional neural network-based diagnosis system for citrus fruit diseases. <italic>Frontiers in Genetics</italic>, 14, 1253934. [<uri>https://doi.org/10.3389/fgene.2023.1253934</uri>] </mixed-citation>
      </ref>
      <ref id="ref012">
        <label>[12]</label>
        <mixed-citation> Lokesh, G. H., Kumar, S., &amp; Reddy, P. (2024). Intelligent plant leaf disease detection using generative adversarial networks: A case-study of cassava leaves. <italic>The Open Agriculture Journal</italic>, 18(1). [<uri>http://dx.doi.org/10.2174/0118743315288623240223072349</uri>] </mixed-citation>
      </ref>
      <ref id="ref013">
        <label>[13]</label>
        <mixed-citation> Wang, C., Zhang, Y., &amp; Li, J. (2020). Identification of vegetable leaf diseases based on improved Multi-scale ResNet. <italic>Transactions of the Chinese Society of Agricultural Engineering</italic>, 36, 209-217. </mixed-citation>
      </ref>
      <ref id="ref014">
        <label>[14]</label>
        <mixed-citation> Ren, S., Zhang, Y., &amp; Li, X. (2020). Recognition and segmentation model of tomato leaf diseases based on deconvolution-guiding. <italic>Chinese Journal of Agricultural Engineering</italic>, 36(12), 186-195. </mixed-citation>
      </ref>
      <ref id="ref015">
        <label>[15]</label>
        <mixed-citation> Guo, X., Fan, T., &amp; Shu, X. (2019). Tomato leaf diseases recognition based on improved multi-scale AlexNet. <italic>Transactions of the Chinese Society of Agricultural Engineering</italic>, 35(13), 162-169. </mixed-citation>
      </ref>
      <ref id="ref016">
        <label>[16]</label>
        <mixed-citation> Qiaoli, Z., Zhang, Y., &amp; Li, J. (2022). Identification of tomato leaf diseases based on improved lightweight convolutional neural networks MobileNetV3. <italic>Smart Agriculture</italic>, 4(1), 47. [<uri>https://doi.org/10.12133/j.smartag.SA202202003</uri>] </mixed-citation>
      </ref>
      <ref id="ref017">
        <label>[17]</label>
        <mixed-citation> Khalid, M., Khan, S., &amp; Ali, A. (2023). Real-time plant health detection using deep convolutional neural networks. <italic>Agriculture</italic>, 13(2), 510. [<uri>https://doi.org/10.3390/agriculture13020510</uri>] </mixed-citation>
      </ref>
      <ref id="ref018">
        <label>[18]</label>
        <mixed-citation> Panchal, A. V., Patel, S., &amp; Shah, M. (2023). Image-based plant diseases detection using deep learning. <italic>Materials Today: Proceedings</italic>, 80, 3500-3506. [<uri>https://doi.org/10.1016/j.matpr.2021.07.281</uri>] </mixed-citation>
      </ref>
      <ref id="ref019">
        <label>[19]</label>
        <mixed-citation> Shoaib, M., Shah, B., &amp; Khan, S. (2022). Deep learning-based segmentation and classification of leaf images for detection of tomato plant disease. <italic>Frontiers in Plant Science</italic>, 13, 1031748. [<uri>https://doi.org/10.3389/fpls.2022.1031748</uri>] </mixed-citation>
      </ref>
      <ref id="ref020">
        <label>[20]</label>
        <mixed-citation> Islam, M. M., Rahman, M. M., &amp; Hossain, M. A. (2022). BdSLW-11: Dataset of Bangladeshi sign language words for recognizing 11 daily useful BdSL words. <italic>Data in Brief</italic>, 45, 108747. [<uri>https://doi.org/10.1016/j.dib.2022.108747</uri>] </mixed-citation>
      </ref>
      <ref id="ref021">
        <label>[21]</label>
        <mixed-citation> Xu, X., Li, W., &amp; Duan, Q. (2021). Transfer learning and SE-ResNet152 networks-based for small-scale unbalanced fish species identification. <italic>Computers and Electronics in Agriculture</italic>, 180, 105878. [<uri>https://doi.org/10.1016/j.compag.2021.105878</uri>] </mixed-citation>
      </ref>
      <ref id="ref022">
        <label>[22]</label>
        <mixed-citation> Bansal, M., Kumar, S., &amp; Sharma, A. (2023). Transfer learning for image classification using VGG19: Caltech-101 image data set. <italic>Journal of Ambient Intelligence and Humanized Computing</italic>, 1-12. [<uri>https://doi.org/10.1007/s12652-021-03488-z</uri>] </mixed-citation>
      </ref>
      <ref id="ref023">
        <label>[23]</label>
        <mixed-citation> Li, Y., Zhang, Y., &amp; Wang, C. (2023). MobileNetV3-CenterNet: A target recognition method for avoiding missed detection effectively based on a lightweight network. <italic>Journal of Beijing Institute of Technology</italic>, 32(1), 82-94. [<uri>https://dx.doi.org/10.15918/j.jbit1004-0579.2022.076</uri>] </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>
