EP4107656A2 - Quantifizieren des pflanzenbefalls durch schätzen der anzahl der biologischen objekte auf den blättern mit faltungsneuronalen netzen, die trainingsbilder verwenden, die durch einen halbüberwachten ansatz erhalten wurden - Google Patents
Quantifizieren des pflanzenbefalls durch schätzen der anzahl der biologischen objekte auf den blättern mit faltungsneuronalen netzen, die trainingsbilder verwenden, die durch einen halbüberwachten ansatz erhalten wurdenInfo
- Publication number
- EP4107656A2 EP4107656A2 EP21705986.4A EP21705986A EP4107656A2 EP 4107656 A2 EP4107656 A2 EP 4107656A2 EP 21705986 A EP21705986 A EP 21705986A EP 4107656 A2 EP4107656 A2 EP 4107656A2
- Authority
- EP
- European Patent Office
- Prior art keywords
- leaf
- images
- plant
- image
- color
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Granted
Links
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F18/00—Pattern recognition
- G06F18/20—Analysing
- G06F18/24—Classification techniques
- G06F18/241—Classification techniques relating to the classification model, e.g. parametric or non-parametric approaches
- G06F18/2413—Classification techniques relating to the classification model, e.g. parametric or non-parametric approaches based on distances to training or reference patterns
- G06F18/24133—Distances to prototypes
- G06F18/24137—Distances to cluster centroïds
- G06F18/2414—Smoothing the distance, e.g. radial basis function networks [RBFN]
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06N—COMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computing arrangements based on biological models
- G06N3/02—Neural networks
- G06N3/08—Learning methods
- G06N3/082—Learning methods modifying the architecture, e.g. adding, deleting or silencing nodes or connections
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06N—COMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computing arrangements based on biological models
- G06N3/02—Neural networks
- G06N3/04—Architecture, e.g. interconnection topology
- G06N3/045—Combinations of networks
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06N—COMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computing arrangements based on biological models
- G06N3/02—Neural networks
- G06N3/04—Architecture, e.g. interconnection topology
- G06N3/045—Combinations of networks
- G06N3/0455—Auto-encoder networks; Encoder-decoder networks
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06N—COMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computing arrangements based on biological models
- G06N3/02—Neural networks
- G06N3/04—Architecture, e.g. interconnection topology
- G06N3/0464—Convolutional networks [CNN, ConvNet]
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06N—COMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computing arrangements based on biological models
- G06N3/02—Neural networks
- G06N3/08—Learning methods
- G06N3/09—Supervised learning
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T7/00—Image analysis
- G06T7/0002—Inspection of images, e.g. flaw detection
- G06T7/0012—Biomedical image inspection
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T7/00—Image analysis
- G06T7/10—Segmentation; Edge detection
- G06T7/11—Region-based segmentation
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V10/00—Arrangements for image or video recognition or understanding
- G06V10/40—Extraction of image or video features
- G06V10/44—Local feature extraction by analysis of parts of the pattern, e.g. by detecting edges, contours, loops, corners, strokes or intersections; Connectivity analysis, e.g. of connected components
- G06V10/443—Local feature extraction by analysis of parts of the pattern, e.g. by detecting edges, contours, loops, corners, strokes or intersections; Connectivity analysis, e.g. of connected components by matching or filtering
- G06V10/449—Biologically inspired filters, e.g. difference of Gaussians [DoG] or Gabor filters
- G06V10/451—Biologically inspired filters, e.g. difference of Gaussians [DoG] or Gabor filters with interaction between the filter responses, e.g. cortical complex cells
- G06V10/454—Integrating the filters into a hierarchical structure, e.g. convolutional neural networks [CNN]
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V10/00—Arrangements for image or video recognition or understanding
- G06V10/40—Extraction of image or video features
- G06V10/56—Extraction of image or video features relating to colour
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V10/00—Arrangements for image or video recognition or understanding
- G06V10/70—Arrangements for image or video recognition or understanding using pattern recognition or machine learning
- G06V10/762—Arrangements for image or video recognition or understanding using pattern recognition or machine learning using clustering, e.g. of similar faces in social networks
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V10/00—Arrangements for image or video recognition or understanding
- G06V10/70—Arrangements for image or video recognition or understanding using pattern recognition or machine learning
- G06V10/764—Arrangements for image or video recognition or understanding using pattern recognition or machine learning using classification, e.g. of video objects
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V10/00—Arrangements for image or video recognition or understanding
- G06V10/70—Arrangements for image or video recognition or understanding using pattern recognition or machine learning
- G06V10/82—Arrangements for image or video recognition or understanding using pattern recognition or machine learning using neural networks
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V20/00—Scenes; Scene-specific elements
- G06V20/10—Terrestrial scenes
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V20/00—Scenes; Scene-specific elements
- G06V20/10—Terrestrial scenes
- G06V20/188—Vegetation
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06N—COMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N20/00—Machine learning
- G06N20/10—Machine learning using kernel methods, e.g. support vector machines [SVM]
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06N—COMPUTING ARRANGEMENTS BASED ON SPECIFIC COMPUTATIONAL MODELS
- G06N3/00—Computing arrangements based on biological models
- G06N3/02—Neural networks
- G06N3/04—Architecture, e.g. interconnection topology
- G06N3/048—Activation functions
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T2207/00—Indexing scheme for image analysis or image enhancement
- G06T2207/20—Special algorithmic details
- G06T2207/20081—Training; Learning
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T2207/00—Indexing scheme for image analysis or image enhancement
- G06T2207/20—Special algorithmic details
- G06T2207/20084—Artificial neural networks [ANN]
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06V—IMAGE OR VIDEO RECOGNITION OR UNDERSTANDING
- G06V20/00—Scenes; Scene-specific elements
- G06V20/10—Terrestrial scenes
- G06V20/17—Terrestrial scenes taken from planes or by drones
Definitions
- the disclosure generally relates to image processing by computers, and more in particular relates to techniques for quantifying plant infestation by estimating the number of biological objects such as insects on plant leaves.
- insects interact with the plant (for example by consuming part of the leaves).
- the insects can cause diseases or other abnormal conditions of the plant.
- the plant does not survive the presence of the insects.
- the plants should become food (i.e. crop for humans or animals), and insects being present on leaves are not desired at all. Food security is of vital importance.
- annotations identify the location of insects by coordinates. Annotations can also differentiate between insect species and growing stages of the insects.
- a computer receives annotations from human expert users who inspect the images on a display. Users interact with the computer in variations, known as inter rater variability (several persons) and intra-rater variability (same person in different moments). Zooming in and out is a further source of possible errors and variations.
- a computer-implemented method for generating a training set with annotated images provide improvements.
- the annotated images are to be used to train a convolutional neural network (CNN) for quantifying plant infestation.
- the CNN estimates the number of insects on leaves of plants.
- the computer receives leaf-images showing leaves and showing insects on the leaves.
- the leaf-images are coded in a first color-coding.
- the computer changes the color coding of the pixels of the leaf-images to a second color-coding.
- the contrast between the pixels for the insects and the pixels for the leaves is higher in the second color-coding than in the first color-coding.
- the computer assigns pixels in the second color-coding to a first binary value or to a second binary value.
- the computer differentiates - in the leaf-images with binary coding - areas with contiguous pixels in the first binary value into into non-insect areas and insect areas by an area size criterion.
- the computer identifies pixel-coordinates of the insect areas, wherein the pixel-coordinates identify rectangular tile-areas with insects in the center.
- the computer annotates the leaf-images in first color-coding by assigning the pixel-coordinates to corresponding tile-areas and thereby obtains the annotated image.
- the computer receives the leaf-images as images that show isolated segmented leaves.
- the computer receives the images with isolated segmented leaves by performing leaf segmentation with a CNN that has been trained by processing leaf- annotated plant images.
- the computer receives the first color-coding as RGB-coding and performs changing the color-coding with a transformation from RGB-coding to XYZ-coding.
- changing the color-coding further comprises to obtain the second color-coding by disregarding the Z-component from the XYZ-coding.
- the computer assigns the binary values by clustering the pixels into color clusters.
- the computer identifies the color clusters by a support-vector machine.
- the computer identifies the pixel- coordinates as squares with 96 x 96 pixels each.
- the computer further classifies the tiles according to insect classes and thereby sorts out false positives.
- the computer applies convolutional neural networks that had been trained previously in a training phase.
- the production phase is summarized first: [0020]
- the computer receives a plant-image taken from a particular plant.
- the plant- image shows at least one of the leaves of the particular plant, the so-called main leaf.
- the computer uses a first convolutional neural network to process the plant-image to derive a leaf- image being a contiguous set of pixels that show a main leaf of the particular plant completely (i.e. as a whole).
- the first convolutional neural network has been trained by a plurality of leaf- annotated plant-images, wherein the plant-images had been annotated to identify main leaves.
- the computer splits the leaf-image into a plurality of tiles.
- the tiles are segments or portions of the plant-image having pre-defined tile dimensions.
- the computer uses a second convolutional neural network to separately process the plurality of tiles to obtain a plurality of density maps having map dimensions that correspond to the tile dimensions.
- a training phase the network has been trained by processing insect-annotated plant-images.
- a first subset of insect-annotated plant-images is obtained by interacting with an expert user (human-made annotations), and a second subset insect-annotated plant-images is obtained by performing the mentioned method for generating a training set (computer-made annotations).
- the training comprised the calculation of convolutions for each pixel based on a kernel function, leading to density maps.
- the density maps have different integral values for tiles showing insects and tiles not showing insects.
- the computer combines the plurality of density maps to a combined density map in the dimension of the leaf-image, and integrates the pixel values of the combined density map to an estimated number of insects for the main leaf.
- FIG. 1 illustrates an overview to a computer-implemented approach to teach - in a training phase - convolutional neural networks (CNN) to count insects on leaves
- FIG. 2 illustrates an overview to a computer-implemented approach to count insects on leaves during a production phase
- FIG. 3 illustrates an overview of computer-implemented methods
- FIG. 4 illustrates a diagram with insect development stages, insect species, and counting classes
- FIG. 5 illustrates a plant-image showing a particular plant with leaves and insects
- FIG. 6 illustrates a user interface of a computer as well as an expert user who annotates images
- FIG. 7 illustrates a leaf-image
- FIG. 8 illustrates an image being split into tiles or sub-regions
- FIG. 9 illustrates a CNN with a splitter module at its input and with a combiner module at its output;
- FIG. 10 illustrates a CNN with layers, in a general overview
- FIGS. 11A and 11B illustrate a pixel value filter that is implemented as part of a CNN layer in implementations that differentiate more than two insect classes
- FIG. 12 illustrates a CNN with multiple output channels
- FIG. 13 illustrates a set of insect-annotated images, with sub-sets for training, for validating and for testing;
- FIG. 14 illustrates pixel-by-pixel color-coding and illustrates the concept of assigning colors to binary pixel values
- FIG. 15 illustrates pixel-by-pixel color-coding and assigning colors to binary pixel values, with changing the color-coding.
- FIG. 16 illustrates clustering as a tool for assigning colors to binary pixel values
- FIG. 17 illustrates a sequence of images
- FIG. 18 illustrates a method-flow chart of computer-implemented method for generating a training set with annotated images
- FIG. 19 illustrates a CNN being trained with human-annotated images and with computer-annotated images
- FIG. 20 illustrates a generic computer system.
- image stands for the data-structure of a digital photograph (i.e., a data- structure using a file format such as JPEG, TIFF, BMP, RAW or the like).
- take an image stands for the action of directing a camera to an object (such as a plant) and letting the camera store the image.
- insect is a noun (in usual language) that most readers can easily apply for counting. The skilled person reading "one insect” or “two insects” immediately understands.
- insect is also used to represent biological objects that are located on parts of the plant be counted (i.e., on the leaves). A biological object (to be counted) has a physical size that is relatively smaller than the part on that it is located. It is also noted that the objects are located on one part. Since the plant images are processed to images showing one part (such as one leaf, by segmentation), the size relation also transfers to the image.
- the biological objects can be insects or can be arachnids (i.e., being arthropoda), or the biological objects can be mollusca (not arthropod).
- the internal structure of the biological objects does not matter, as long as it fits the size criterion.
- the object has an exoskeleton, a segmented body, and paired jointed appendages (as arthropods have) or not.
- the different number of legs e.g., insects 6 legs, arachnids 8 legs, or even no legs as with snails
- the objects can also be spots on the surface of the plant parts (spots that are the result of biological processes, such as fungi interacting with the plant or the like, animal excrements, etc.).
- suitable measures e.g., countermeasures.
- the use of the term "insects" is applicable to phrases such as "insect-annotated” or the like to the meaning "object -annotated”.
- the interaction of the biological object with the plant does not matter.
- the biological object can be pest or beneficial.
- annotation stands for meta-data that accompanies an image to identify particular properties regarding the content (of the image).
- an annotation can identify the border or edge of a plant leaf (i.e., "leaf-annotation”), or can identify presence, type, location etc. of an insect sitting on that leaf (i.e., "insect-annotation”)
- annotated image indicates the availability of such meta-data for an image (or for a sub-region of that image), but there is no need to store the meta-data and the image in the same data-structure.
- drawings illustrate annotations as part of an image, such as by polygon and/or dots, but again: the annotations are meta-data and there is no need to embed them into the data-structure of the image.
- a human-made insect annotation is obtained through an expert user who looks at the display of an image and interacts with a computer (resulting in a "insect-by-human-annotated images” or short "human-annotated images”).
- a computer-made annotation is obtained by a computer that performs a method for providing annotated images (cf. FIGS. 14-18, "insect-by-computer-annotated images", or "computer-annotated images”).
- insects stands for animalia in the phylum “Arthropoda” or 1ARTHP (EPPO-code by the European and Mediterranean Plant Protection Organization).
- the insects are of the subphylum “Hexapoda” (1HEXAQ).
- stage (also “development stage”, “growing stage”) identifies differences in the life-cycle (or metamorphosis) of insects, wherein an insect in a first stage has a different visual appearance than an insect in a second, subsequent stage.
- Biologists can differentiate the stages (or “stadia”) by terms such as egg, larva, pupa, and imago. Other conventions can also be used, such as “adults” and “nymphs”, or even standardized numerical identifiers such as “nln2", “n3n4" and so on. Development stages of the plants are not differentiated.
- count is short for "estimating a number”, such as for estimating the number of insects on a leaf.
- train as a label for a first process - the "training process” - that enables CNNs to count insects, and for the particular task to train a particular CNN by using annotated images.
- the description refers to hardware components (such as computers, cameras, mobile devices, communication networks) in singular terms.
- implementations can use multiple components.
- the camera taking a plurality of images comprises scenarios in that multiple cameras participate so that some images are taken from a first camera, some image are taken from a second camera and so on.
- the description provides an overview to the application of CNNs in two phases, and thereby introduces pre-processing activities as well as computer- implemented methods.
- the description introduces insects in different development stages to be counted.
- the description investigates details (regarding plants, images, image portions to be processed). With FIG. 13, the description discusses accuracy.
- FIGS. 1-3 illustrate overviews to computer-implemented approaches as in FIG. 1, to train convolutional neural networks (CNN) to count insects on leaves in a training phase **1, and as in FIG. 2, to quantify infestation by actually counting insects on leaves during a (subsequent) production phase **2 (FIG. 2).
- CNN convolutional neural networks
- FIGS. 1-2 illustrate plants 111/112 (with leaves and insects), cameras 311/312 to take plant-images 411/412, and computers 201/202 with CNNs to perform computer-implemented methods 601B/602B/701B/702B.
- the figures also illustrate human users 191/192.
- FIGS. 1-2 illustrate computers 201/202 by rectangles with bold frames.
- Computers 201/202 implement methods 601B, 602B, 701B and 702B (FIG. 3) by techniques that are based on Machine Learning (ML).
- FIGS. 1-2 also illustrate computer 301 and mobile device 302, performing auxiliary activities (or participating therein), such as taking images, transmitting images, receiving annotations, and forwarding results to other computers, such as estimation values.
- auxiliary activities or participating therein
- auxiliary activities are pre-processing activities that prepare method executions.
- the pre-processing activities are illustrated by references 601A, 602A, 701A and 702A.
- Computers 201/202 use CNNs and other modules to be explained below (such as user interfaces, databases, splitter and combiner modules etc.). While FIGS. 1-2 just introduce the CNNs, the other figures provide details for pre-processing images and for setting parameters to the CNNs. CNNs 261 and 271 are being trained in the training phase **1 to become trained CNNs 272 and 272, respectively. In other words, the difference between untrained and trained CNNs is the availability of parameters obtained through training.
- FIG. 3 illustrates an overview to computer-implemented methods 601B, 602B, 701B and 702B.
- the methods are illustrated in a matrix.
- counting insects on leaves is divided into a sequence, with simplified: a first sub-sequence (illustrated by the column on the left side) to identify leaves on images, and a second sub-sequence (illustrated by the column on the right side), to count the insects on the identified leaves.
- FIG. 3 differentiates pre-processing activities 601A, 602A, 701A, and 702A (such as taking images and annotating images) from computer-implemented methods 601B, 602B,
- Methods 601B and 602B are performed with CNNs 261/262, and methods 701B and 702B are performed with CNNs 271/272.
- CNNs 261/262 and CNNs 271/272 differ from each other by parameters (explained below).
- the CNNs use density map estimation techniques, where - simplified - the integral of the pixel values leads to the estimated insect numbers. In other words, counting is performed by calculating an integral.
- the estimated numbers NEST can be non-integer numbers. For the above-mentioned purpose (to identify appropriate countermeasures against the infestation), the accuracy of NEST is sufficient.
- Training phase **1 is illustrated in the first row of FIG. 3 in reference to FIG. 1.
- camera 311 takes a plurality of plant-images
- Computer 301 interacts with expert user 191 to obtain leaf-annotations and to obtain insect-annotations.
- User 191 can have different roles (details in FIG. 6).
- Combinations of images and annotations are provided as annotated images 461, 471.
- the description differentiates leaf-annotated plant-images 461 and insect-annotated leaf-images 471. It is however noted that a particular image can have both leaf- annotations and insect- annotations.
- Method 800 for providing a training set (with annotated image).
- Method 800 bypasses the human user and generate computer-made annotations (i.e. images 473).
- Computer 301 forwards annotated images 461, 471 to computer 201.
- computer 303 forwards images 473 to computer 201.
- computer 201 receives the plurality of plant-images in combination with the leaf-annotations (collectively "leaf-annotated plant-images").
- Computer 201 uses a sub-set of the plurality and trains CNN 261 to identify a particular leaf in a plant-image (that is not annotated). Thereby, computer 201 converts un-trained CNN 261 into trained CNN 262. In other words, CNN 262 is the output of computer 201.
- computer 201 receives the plurality of leaf-annotated plant-images in combination with insect-annotations (collectively "insect-annotated leaf- images").
- Computer 201 then trains CNN 271 to count insects on particular leaves. Thereby, computer 201 turns un-trained CNN 271 into trained CNN 272. In other words, CNN 272 is output of computer 201 as well.
- computer 201 can also perform method 701B in two variations. In the first variation, computer 201 trains CNN 271 with human-annotated images 471 and in the second variation, computer 201 trains CNN 271 with computer-annotated images 473. Training uses the results of the both variations, but in terms of the loss function, computer 201 puts more weight on human-annotated images 471. [0078] It is noted that the description assumes the annotations to be made for the same plurality of plant-images 411. This is convenient, but not required.
- the pluralities can be different.
- the plurality of plant-images 411 to be leaf-annotated can show non- infested plants.
- leaf-annotated plant-images 471 from such healthy plants
- Providing insect-annotations could be performed for images that are not segmented to leaves.
- Production phase **2 is illustrated in the second row of FIG. 3, in reference to FIG. 2.
- camera 312 of device 302 takes plant-image 412 and forwards it to computer 202.
- computer 202 In performing method 602B, computer 202 (FIG. 2) uses CNN 262 to identify a particular leaf (and thereby creates a leaf-image). Subsequently, in performing method 702B, computer 202 uses CNN 272 and processes leaf-image 422. Thereby, computer 202 counts insects (and potentially other objects if trained accordingly) on that particular leaf. Thereby, computer 202 obtains the estimated number of insects per leaf NEST as the result.
- Plant 111 has leaves 121, and leaves 121 are occupied by insects 131 and non-insect objects 141.
- Camera 311 takes (a plurality of) plant- images 411 that are processed in computer 201 during training phase **1.
- Training phase **1 has two sub-phases.
- the first sub-phase comprises activities such as taking the plurality of plant-images 411, letting expert user 191 provide annotations to plant-images 411 to computer 301.
- Computer 301 receiving annotations results in annotated images 461, 471 that are leaf-annotated images and insect-annotated images (FIG. 6).
- Receiving annotations can also be considered as the supervised learning part.
- computer 201 uses annotated images 461, 471 to train CNN 261 to identify a main leaf (i.e. a semantic segmentation of the plant image) and to train CNN 271 to identify the insects (i.e., on the main leaf that was identified earlier).
- a main leaf i.e. a semantic segmentation of the plant image
- CNN 271 to identify the insects (i.e., on the main leaf that was identified earlier).
- Persons of skill in the art can apply suitable training settings.
- computer 201 is illustrated by a single box, it can be implemented by separate physical computers.
- the same principle applies for plant 111 and for camera 311.
- the plant and the camera do not have to be the same for all images. It is rather expected to have plant-images 411 from cameras 311 with different properties.
- the plurality of images 411 represents a plurality of plants 111. There is no need for a one-to-one relation, so one particular plant may be represented by multiple images.
- Training the CNNs can be seen as the transition from the training phase to the production phase. As in FIG. 1, CNNs 261 and 271 are being trained to become trained-CNNs 262 and 272.
- Training phase **1 is usually performed once, in supervised learning with expert user 191.
- the setting for the training phase with camera 311 taking plant-images 411 (as reference images), with expert user 191 annotating plant-images 411 (or derivatives thereof) and with computer-implemented processing will be explained.
- the description assumes that training phase **1 has been completed before production phase **2. It is however possible to perform training phase **1 continuously and in parallel to production phase **2.
- FIG. 2 it illustrates an overview of a computer-implemented approach by that a computer - illustrated as computer 202 - counts insects 132 on leaves 122 of plants 112 in an exemplary application in an agricultural field. It does not matter if the field is located in open air or located in a green-house.
- Simplified, computer 202 processes plant-image 412 received from mobile device 302 through communication network 342.
- one image 412 is theoretically enough.
- Leaves 122 are so-called "infested leaves” because insects are located on them. Counting can be differentiated for insects of particular class 132(1) (illustrated by plain ovals) and - optionally - of particular class 132(2) (bold ovals). Optionally, counting can consider further classes (cf. FIGS. 4 and 11).
- Non-insect objects 142 are not necessarily to be counted. Such objects 142 can be located within the leaf and can be structural elements of leaves 122, such as damages on the leaf, shining effects due to light reflection or the like. It is noted that many insects camouflage themselves. Therefore, for the computer it might be difficult to differentiate insects 1S2 and non-insect objects 142.
- Insect classes (1) and (2) are defined by a particular insect species, and (i.e., logical AND) by a particular development stage (cf. FIGS. 2 and 4).
- FIG. 4 A more fine-tuned granularity with more classes is given in FIG. 4, and an adaptation of the CNNs to such granularity is given in FIGS. 11-12.
- the term "insect” is used synonymous to "bug”.
- Plant 112 has a plurality of leaves 122. For simplicity, only two leaves 122-1 and 122-2 are illustrated. Leaves 122 are occupied by insects 1S2 (there is no difference to the training phase **1). For convenience, FIG. 2 is not scaled, with the size of the insects being out of proportion.
- Mobile device 302 can be seen as a combination of an image device (i.e. camera 312), processor and memory. Mobile device 302 is readily available to the farmers, for example as a so-called “smartphone” or as a “tablet”. Of course, mobile device 302 can be regarded as a "computer”. It is noted that mobile device 302 participate in auxiliary activities (cf. FIG. 3,
- Field user 192 tries to catch at least one complete leaf (here leaf 122-1) into (at least one) plant-image 412. In other words, field user 192 just makes a photo of the plant.
- field user 192 may look at user interface 392 (i.e. at the visual user-interface of device 302) that displays the plant that is located in front of camera 312.
- Mobile device 302 then forwards plant-image 412 via communication network 342 to computer 202.
- communication network 342 suggests, computer 202 can be implemented remotely from mobile device 302.
- Computer 202 returns a result that is displayed to user interface 392 (of mobile device 302).
- N(l) 3 insects of class (1) (i.e., insects 132(1))
- N(2) 2 insects of class (2) (i.e., insects 132(2)) counted.
- the numbers N(l), N(2) (or N in general) are numbers-per-leaf, not numbers per plant (here in the example for main leaf 122-1).
- the numbers correspond to N EST (with N EST being rounded to the nearest integer N).
- field user 192 can identify countermeasures to combat infestation with better expectation of success.
- main leaf does not imply any hierarchy with the leaves on the plant, but simply stands for that particular leaf for that the insects are counted.
- Adjacent leaf 122-2 is an example of a leaf that is located close to main leaf 122-1, but for that insects are not to be counted. Although illustrated in singular, plant 112 has one main leaf but multiple adjacent leaves. It is assumed that plant-image 412 represents the main leaf completely, and represents the adjacent leaves only partially. This is convenient for explanation, but not required [00104] It is usual that main leaf 122-1 is on top of adjacent leaf 122-2 (or leaves). They overlap each other and it is difficult to identify the edges between one from the other.
- the numbers N are derived from estimated numbers N EST , the description describes an approach to accurately determine N.
- counting comprises two major sub-processes: first, differentiating main and adjacent leaves (also called leaf identification, segmentation), and second, counting the insects on the main leaf only.
- identifying the main leaf prior to counting keeps the number of "false negatives" and "false positives” negligible.
- the communication between mobile device 302 and computer 202 via communication network 342 can be implemented by techniques that are available and that are commercially offered by communication providers.
- Computer 202 has CNN 262/272 that performs computer-implemented method 602B and 602B (details in connections with FIGS. 10). CNNs 262 and 272 have been trained before (methods 601B and 701B).
- computers 201/202 use operating system (OS) Linux, and the module that executes methods 601B/602B, 701B/702B was implemented by software in the Python programming language. It is convenient to implement the modules by a virtualization with containers. Appropriate software is commercially available, for example, from Docker Inc. (San Francisco, California, US).
- SaaS software-as-a-service
- computer 202 has other modules, for example, a well-known REST API (Representational State Transfer, Application Programming Interface) to implement the communication between mobile device 302 and computer 202 can use.
- Computer 202 appears to mobile device 302 as a web-service. The person of skill in the art can apply other settings.
- the time it takes computer 202 with CNNs 262/272 (performing the method) to obtain NEST depends on the resolution of plant-image 412. Performing methods 602B and 702B may take a couple of seconds. The processing time rises with the resolution of the image. (The processing time has been measured in test runs. For plant-image 412 with 4000 x 6000 pixels, the processing time was approximately 9 seconds.)
- the conditions for catching images are not always ideal. For example, there are variations in the acquisition distance (between camera and plant, users holding the mobile devices at different heights), the illumination (e.g., sunny daylight or rainy/cloudy daylight, some leaves may shadow other leaves), the surface of the plant (e.g., dust or rain-drops on the plant etc.), perspective (e.g., taking images from the side or from the top to name two extremes), focusing (e.g., sharp image for non-relevant parts of the plant), resolution (e.g., mobile devices with 24 M pixel cameras, versus devices with less pixels), and so on.
- the illumination e.g., sunny daylight or rainy/cloudy daylight, some leaves may shadow other leaves
- the surface of the plant e.g., dust or rain-drops on the plant etc.
- perspective e.g., taking images from the side or from the top to name two extremes
- focusing e.g., sharp image for non-relevant parts of the plant
- resolution e.g
- insects 131/132 cf. FIGS. 1-2
- FIGS. 1-2 the description shortly investigates the objects to be counted: insects 131/132 (cf. FIGS. 1-2), but then turns to a discussion of problems with existing technology and of solution approach that is adapted to count insects.
- plants 111/112 are eggplants (Solanum melongena, EPPO- code SOLME) and insects 131/132 are of the species whitefly (Bemisia tabaci EPPO-code BEMITA).
- plants 111/112 are eggplants as well, and insects 131/132 are of the species thrips (Franklinella occidentalis, EPPO-code FRANOC).
- pre-processing 601A and executing method 601B i.e. to train CNN 261/262 to segment leaves
- pre-processing 602A and executing method 602B would be performed with a second plant species.
- the second plant species can belong to other crops such as for example cotton, soy bean, cabbage, maize (i.e., corn).
- insects change appearance in the so-called metamorphosis with a sequence of development stages.
- FIG. 4 illustrates a diagram with development stages (A), (B), (C ) and (D) for insects 131/132 (cf. FIGS. 1-2), insect species (i) and (ii), and counting classes (1), (2), (3), (4)
- the development stages occur in a predefined sequence with state transitions: from stage (A) to (B), from (B) to (C), from (C) to (D).
- the arrows are dashed, just to illustrate that other transitions (such as from (B) to (D)) are possible.
- Biologists can associate the stages with semantics relating to the age of the insects, such as "egg", “nymph”, “adult”, “empty pupae” (an insect has left the pupa and only the skin of the pupa is left), with semantics relating to life and death.
- stages As particular way to express stages is the "nln2"/ "n3n4" nomenclature, well known in the art.
- Details for the appearance in each stage are well-known. Just to mention one point, insects can develop wings. For example, the presence or absence of wings can indicate particular development stage for thrips.
- FIG. 4 illustrates a stage-to-species matrix, with stages (A) to (D) in columns, and insect species (i) and (ii) in rows.
- stages (A) to (D) in columns, and insect species (i) and (ii) in rows.
- insect species i) and (ii) in rows.
- stages (i) and (ii) in rows.
- Insects of both species develop through the (A) to (D) stages (of course separately: (i) do not turn into (ii) or vice versa).
- the black dots at the column/row crossings indicate that insects of particular stage/species combinations should be counted. This is a compromise between accuracy (e.g. infestation critical for black dotted situations, but countermeasures available) and efforts (annotations, calculations, training etc.).
- expert users 191 can annotate insects in the particular stage/species combinations on plant-images (or leaf-images) and the CNN can be trained with such annotations.
- computer 202 with CNN 272 can count the insects accordingly.
- CNN 271/272 is trained to provide NEST as the number of species (i) insects in stages (B) and (C), without differentiating (B) and (C), that is NEST (i)(B)(C) [00134]
- CNN 271/272 is trained to provide N EST in 2 separate numbers (cf.. the introduction in FIGS. 1-2): N EST (ii) (B) and NEST (ii) (C)
- CNN 271/272 is trained to provide NEST in 4 separate numbers: NEST (i) (A), NEST (i) (B), NEST (i) (C), and NEST (i) (D).
- NEST (i) (A) The rectangles are illustrated with class numbers (1) to (4), wherein the classes are just alternative notations. The description will explain adaptations to the CNNs for multi-class use cases (use cases 2 and 3) in connection with FIGS. 11-12.
- the impact of the insects to the plant can be different for each development stage. For example, it may be important to determine the number of nymphs (per leaf), the number of empty pupae and so on. Differentiating between young and old nymphs can indicate the time interval that has passed since the arrival of the insects, with the opportunity to fine-tune the countermeasure. For example, adults may lay eggs (and that should be prevented).
- FIG. 5 illustrates plant-image 411/412 (dashed frame, cf. FIGS. 1-2).
- Plant-image 411/412 shows a particular plant with leaves 421/422 and with insects 431/432.
- FIG. 5 is simplified and uses symbols for the leaves (without illustrating the characteristic leaf shape).
- Leaf 421-1 corresponds to leaf 121-1 (of FIG. 1) and leaf 421-2 corresponds to leaf 121-2 (of FIG. 1), illustrated partly.
- Leaf 422-1 corresponds to leaf 122-1 (of FIG. 1) and the leaf 422-2 corresponds to leaf 122-2 (of FIG. 2), illustrated partly as well.
- FIG. 5 illustrates insects 431/432 by small squares, there are some of them on leaf 421-1/422-1 and some of them on leaf 421-2/422-2.
- Non-insect object 441/442 is symbolized by a small square with round corners.
- the insects can belong to different classes (cf. FIG. 4).
- FIG. 5 illustrates an image (being a data-structure) so that "insect 431/432" actually symbolizes the pixels that show the insect (likewise for 441/442).
- plant-image 411 (cf. FIG. 1) can be taken by a high- resolution camera (e.g., 24 mega pixel) or by a main camera of a mobile device.
- a high- resolution camera e.g. 24 mega pixel
- main camera of a mobile device e.g., a main camera of a mobile device.
- images are taken in pluralities. It is noted that the variety of different cameras can be taken into account when taking images for training.
- plant-image 412 is usually taken by camera 312 of mobile device 302 (cf. FIG. 2).
- FIG. 5 is also convenient to explain constraints that arise from the objects (i.e., plants with leaves and insects) and from insufficiencies of mobile device cameras.
- FIG. 5 illustrates image 411/412 in portrait orientation (height larger than width), this is convenient but not required.
- Image coordinates (X, Y) to identify particular pixels are given for convenience. Image dimensions are discussed in terms of pixels.
- image 411/412 can have 6000 pixels in the Y coordinate, and 4000 pixels in the X coordinate (i.e. 24 Mega Pixels, or "24 M").
- the pixel numbers are the property of the camera sensor and can vary.
- Image 411/412 is usually a three-channel color image, with the color usually coded in the RGB color space (i.e., red, green and blue).
- image 412 does not have to be displayed to field user 192. Also, the field scenario will be explained for a single image 412, but in practice it might be advisable for field user 192 to take a couple of similar images 412.
- Image 412 represents reality (i.e. plant 112, leaves 122, insects 132, non-insect objects 142), but with at least the following further constraints.
- plant 111/112 has multiple leaves at separate physical locations. Therefore in image 411/412, one leaf can overlay other leaves. Or in other words, while in reality (cf. FIG. 1), leaves are separate, leaf 421-1 and leaf 421-2 appear as adjacent leaves (422-1 and 422-2 as well).
- each insect of a particular class has a particular color. This color could be called text-book color, or standard color. For example, as the name suggests, an adult whitefly is white (at least in most parts).
- the image would not properly represent the text-book color.
- the color of the insect has a natural variability
- the illumination can be different (e.g., cloudy sky, sunny sky, shadow and so on)
- Camera 311/312 is not an ideal camera. It does not take such different illumination conditions into account. As a consequence, the images may not properly show the color.
- insects can be relatively tiny in comparison to the leaves. For example, an insect can be smaller than one millimeter in length. In contrast to the emphasis in FIGS. 1-2, a couple of hundred insects may occupy a single leaf easily. The insects are also usually relative tiny things for the human eye to detect. This is in sharp contrast to, for example, a single bee in the petal leaves of a flower.
- insects tend to be present on the leaf in pairs (i.e., two insects), or even in triples (i.e., three insects). So in other words, a 30 x 20 pixel portion of image 411/412 might represent two or more insects.
- the pixel numbers 20 x 30 are exemplary numbers, but it can be assumed that insects 431/432 are dimensioned with two-digit pixel numbers (i.e. up to 99 pixels in each of the two coordinates). The same limitations can be true for non-insect objects 441/442.
- CNNs 271/272 use density map estimation (instead of the above-mentioned traditional object detection).
- density maps insects would be represented as areas, and the integral of the pixel values of the area would be approximately 1 (assuming that the pixel values are real numbers, normalized between 0 and 1, and also assuming to have one insect per map). It is noted that for situations in that two insects are located close together and overlapping on the image, there would be a single area, but the sum of the pixel values would be approximately 2.
- insects of two or more stages can be available on a single leaf at the same time. It is a constraint that the differences between two stages can be subtle. For example, on a leaf in reality, insects in stages (C) and (D) may look similar.
- a computer using a conventional computer-vision technique may not recognize the differences.
- an expert user can see differences (on images), and training images can be properly annotated (cf. use case 1).
- the farmer i.e. the user of the mobile device
- the time interval from taking the image to determining the insect-number-per-leaf must be negligible so that the insects do not substantially grow (and/or reproduce, and/or eventually change progress to the next development stage) during that time interval, the insects do no fly away (because the measures are applied to the infested plants).
- the identification and the application of the countermeasures can only start when the insect-number-per-lead has been established.
- a countermeasure - although properly identified - may be applied too late to be effective.
- a countermeasure that is specialized to destroy eggs would not have any effect if the insects have already hatched from the eggs (cf. stage specific countermeasures).
- FIG. 6 illustrates a user interface of computer 301 (cf. FIG. 1).
- FIG. 6 also illustrates expert user 191 annotating images.
- FIG. 6 explains some of the pre-processing activities (cf. FIG. 3, 601A and 701A) in training phase **1.
- FIG. 6 is related to supervised learning.
- FIG. 6 gives more details how to obtain annotated images 461, 471 introduced above in connection with FIGS. 1-3.
- the coordinate system (X, Y) is given for convenience (cf. FIG. 5).
- Expert user 191 conveys ground truth information to the images, not only regarding the presence or absence of a main leaf (by the leaf-annotations), or the presence or absences of particular insects (by the insect-annotations), but also information regarding the position of the main leaf and of the insect in terms of (X, Y) coordinates.
- the annotations can also identify insect species, development stages and so on.
- FIG. 6 illustrates single image 411, annotating is repeated, for example for 1.947 images (leaf annotation). In view of that number, it is noted that expert user 191 is not necessarily always the same person.
- expert user 191 annotates plant-image 411 to obtain leaf-annotated plant-image 461.
- the leaf-annotation identifies the leaf border of the main leaf 421-1 in difference to adjacent leaf 421-2.
- the leaf-annotation can also identify the border between leaf and background (or soil, if visible on the image).
- leaf-annotated plant images show annotated borders between leaf and background, and between leaf and leaf
- user 191 can draw polygon 451 (dashed line) around that part of plant-image 411 that shows the complete leaf (i.e. the main leaf).
- image 411 shows leaf 420-1 as the complete leaf, and shows leaf 420-2 only partially, cf. FIG. 5.
- Computer 301 can close polygon 451 automatically.
- the person of skill in the art can use other user interfaces, for example picture processing tools to manipulate images, for example, by "erasing" the pixels surrounding the main leaf.
- the leaf-annotation allows computer 201 (cf. FIG. 1) for each pixel of plant-image 411 to differentiate if the pixel belongs to the main leaf or not. This differentiation is relevant for performing method 601B (cf. FIG. 3, leaf segmentation). [00183] For the leaf-annotation, it does not matter if the leaf shows insects (or non-insect objects).
- the leaf-annotation allows the CNN being trained to differentiate image regions that show two types of "borders”: between leaf and leaf, and between leaf and background (such as soil). Once trained, the CNN can but the leaf along such borders. In other words, the borders (or margins) stand for a cutting line. Insect annotation
- insect-annotated leaf-image 471 As illustrated on the right side of FIG. 6, user 191 also annotates leaf-image 421 (cf. FIG. 7) to obtain insect-annotated leaf-image 471.
- the insect annotation identifies insects and - optionally - identifies insect classes (cf. species and/or stages, as explained by the classes in FIG. 4).
- the term "insect-annotated” is simplified: image 471 can comprise annotations for the non-insect objects are well.
- the insect annotation also identifies the position of the insects (and/or non-insect objects) by coordinates.
- annotations can take the use cases (cf. FIG. 4) into account.
- the annotations are illustrated by dots with references a to z, with - for example annotation a pointing to a whitefly (i) in stage (C); annotation b pointing to a whitefly (i) in stage (B); annotations y and d pointing to whiteflies (i) in stage (B); annotation e pointing to a non-insect object, being an optional annotation; and annotation z pointing to thrips (ii) in stage (B), but differentiating stages would also be possible [00188] Expert user 191 can actually set the dots next (or above) to the insects.
- a single dot points to a particular single pixel (the "dot pixel” or "annotation pixel”).
- the coordinate of that single pixel at position coordinate (C', U') of an insect (or non-insect object) is communicated to computer 301.
- FIG. 6 illustrates the position coordinate for annotation b, by way of example.
- the user interface can display the dot by multiple pixels, but the position is recorded by the coordinates at pixel accuracy.
- Computer 301 stores the position coordinates as part of the annotation. Coordinates (C', U') can be regarded as annotation coordinates, and the computer would also store the semantic, such as (i)(C) in annotation 1, as (i)(B) in annotation b and so on.
- the insect-annotation (for a particular image 411) is used by computer 201 in training CNN 271, for example, by letting the computer convolute images (i.e., tiles of images) with kernel functions that are centered at the position coordinate (C', U'). Also, the insect-annotation comprises ground truth data regarding the number of insects (optionally in the granularity of the use cases of FIG. 4).
- the annotations can be embedded in an annotated image by dots (in color-coding, e.g. red for stage (C), stage (D), or as X, Y coordinates separately.
- FIG. 7 illustrates leaf-image 421/422.
- Leaf-image 421/422 shows a particular plant with its main leaf and with insects. In difference to plant-image 411/412, leaf-image 421/422 only shows the main leaf, but not the adjacent leaves. The main leaf is the object of interest.
- Leaf-image 421/422 shows the leaf substantially completely (because the insects are to be counted per leaf), and therefore the margin of the leaf is shown substantially completely as well. In that sense, leaf-image 421/422 is a cropped image derived from plant-image 411/412. The image is cropped to the leaf.
- leaf-image 421 and 422 can differ, as described in the following: [00195] In training phase **1, computer 201 obtains leaf-image 421 through interaction with expert user 191, as explained below (cf. FIG. 6, left side). Leaf-image 421 can be considered as the portion of leaf-annotated plant-image 461 that shows the leaf. Leaf-image 421 is illustrated here for explanation only. As explained in connection with FIG. 3, computer 201 processes the plurality of leaf-annotated plant-images 461 to obtain CNN 262, this process uses the annotations.
- computer 202 obtains leaf-image 422 through segmenting plant-image 412 by using (trained) CNN 262 (in method 602B). In the production phase, annotations are not available. It is noted that leaf-image 422 (production phase) is not the same as leaf-image 421 (training phase).
- Reference 429 illustrates portions of leaf-image 421/422 that do not show the main leaf.
- the pixels in portions 429 can be ignored in subsequence processing steps. For example, a processing step by that an image is split into tiles does not have to be performed for portions 429 (because insects are not to be counted according to the insects per leaf definition).
- these portions 429 can be represented by pixels having a particular color or otherwise. In illustrations (or optionally in displaying portions 429 to users), the portions can be for example displayed in black or white or other single-color (e.g., white as in FIG. 7).
- FIG. 8 illustrates image 401/402 being split into tiles 401-k/402-k (or sub-regions).
- the image can be plant-image 411 (training phase), annotated as image 461 or as image 471, leading to tiles 401-k, or plant-image 412 (production phase), leading to tiles 402-k.
- the number of tiles 401-k/402-k in image 401/402 is given be reference K.
- image 400 can have an image dimension of 4000 x 6000 pixels (annotations do not change the dimension).
- the tiles have tile dimensions that are smaller than the image dimensions.
- the tile dimension is 256 x 256 pixels.
- the tile dimensions correspond to the dimension of the input layer of the CNNs (cf. FIG. 10).
- tile 401-k was split out from annotated image 471.
- tile 401-k comprises the pixel with that particular coordinate and tile 401-k takes over this annotation (cf. the dot symbol, with position coordinates (C', U') cf. FIG. 6).
- the person of skill in the art can consider the different coordinate bases (cf. FIG. 6 for the complete image, FIG. 8 for a tile only).
- CNN 271 would learn parameters to obtain density map 501-k with the integral summing up to 1 (corresponding to 1 insect, assuming normalization of the pixel values in the density maps). For example, CNN 271 would take the position coordinate (C', U') to be the center for applying a kernel function to all pixels of tile 401-k.
- tile 402-k was split from image 412 (production phase), it shows insect 432. Of course, the insect is not necessarily at the same position as in "annotated" tile 401-k above in alpha).
- CNN 272 Using the learned parameters, CNN 272 would arrive at density map 502-k with integral 1.
- the example gamma is a variation of the example alpha. Tile 401-k was split up, and annotations are de facto available as well. Although expert user 192 did not provide annotations (dots or the like), the meta-data indicates the absence of an insect.
- CNN 271 would learn parameters to obtain density map 501-k with the integral summing up to 0.
- the example delta is a variation of case beta.
- a non-insect tile 402-k is processed in the production phase (by CNN 272) and it would arrive at a density map with integral 0.
- the person of skill in the art can implement this, for example, by operating CNN 272 in K repetitions (i.e. one run per tile), and combiner module 282 can reconstruct the density maps of the tiles in the same order as splitter module 242 has split them (cf. FIG. 9 for the modules).
- FIG. 9 illustrates CNN 271/272 with splitter module 241/242 and with combiner module 281/282.
- Splitter module 241/242 receives images 411/ 412 (images 411 with annotations as images 461, 471) and provides tiles 401-k/402-k. As it will be explained, the CNNs provide density maps, combiner module 281/282 receives maps 501-k/502-k and provides combined density map 555.
- combiner module 281/282 can compose the image (or combined density map) by overlapping likewise. Pixel values at the same particular coordinate (X, Y coordinates for the 4000 x 6000 pixels) are counted only once.
- Combiner module 281/282 can also calculate the overall integral of the pixel values (of the combined density map), thus resulting in N EST.
- FIG. 10 illustrates CNNs 261/262/271/272 with layers, in a general overview.
- the CNNs are implemented by collections of program routines being executed by a computer such as by computer 201/202.
- FIG. 10 illustrates the CNNs with the input to an input layer and with the output from an output layer.
- FIG. 10 also illustrates (at least symbolically) intermediate layers.
- CNNs 261/262/271/272 are deep networks because they have multiple intermediate layers. The intermediate layers are hidden. In other words, deep learning is applied here.
- FIG. 10 also illustrates some parameters and illustrates intermediate images (being tiles and maps). Since CNNs are well known in the art, the description focuses on the parameters that are applied specially for segmenting by CNNs 261/262 and for counting by CNNs 271/272.
- CNNs 261/271 receive annotated images 461, 471 and turn un-trained CNN 261 into trained CNN 262 (using the leaf-annotated plant-images) and turn un trained CNN 271 into trained CNN 272 (using insect-annotated leaf-images).
- CNNs 262 and 272 receive plant-image 412 and provide output, such as NEST (i.e. the number of insects per leave).
- NEST i.e. the number of insects per leave.
- the CNNs do not receive the images in the original image dimension (e.g., 4000 x 6000 pixels) but in tile/map dimensions (e.g., 224 x 224 pixels).
- FIG. 10 illustrates an example for intermediate data by intermediate images: tile 401/402-k and density map 501-k/502-k. Index k is the tile index explained with FIGS. 8-9.
- FIG. 10 illustrates tile 402-k that shows two insects.
- Tile 402-k is a portion of a plant-image or a portion of a leaf-image and has tile dimensions optimized for processing by the CNN layers.
- tile 402-k has 256 x 256 pixels (or 224 x 224 pixels in a different example).
- Tile 402-k is obtained by splitting an image (details in connection with FIG. 8) to tiles.
- Map 502-k is a density map derived from tile 402-k. Map 502-k has the same dimension as the tile 402-k. In other words, the map dimensions and the tile dimensions are corresponding to each other.
- the density map can be understood as a collection of single-color pixels in X-Y-coordinates, each having a numerical value V(X, Y).
- the integral of the values V of all X-Y-coordinates corresponds to the number of objects (i.e. insects). In the example, the integral is 2 (in an ideal case), corresponding to the number of insects (e.g., two insects shown in tile 402-k).
- map 502-k is obtained by prediction (with the parameters obtained during training).
- one of the processing steps is the application of a kernel function (e.g., a Gaussian kernel) with the kernel center corresponding to an annotation coordinate (C', U'), if an annotation (for an insect) is available in the particular tile 402-k.
- a kernel function e.g., a Gaussian kernel
- C', U' annotation coordinate
- kernel functions are not applied.
- Converting tile 401-k to map 501-k is based on layer-specific parameters obtained by training (i.e., training CNN 271 to become CNN 272). Since the insect-annotations (cf. FIG. 6) indicate the presence (or absence) of an insect (or more insects as here), the annotations are also applicable to the (plurality of tiles). There are tiles with annotations (insects are present) and there are tiles without annotations (insects are not present).
- tile 401-k has the annotation "2 insects". It is noted that both insects can belong to different classes (cf. FIG. 4), the differentiation between classes (i.e. counting the insects in a class-specific approach) is explained in connection with class branching (cf. FIGS. 11-12).
- Networks are publicly available in a variety of implementations, and the networks are configured by configuration parameters.
- Exemplary networks comprise the following network types (or "architectures"): [00231] (i) The UNet type is disclosed by Ronneberger, O., Fischer, P., Brox, T., 2015. U-net:
- FCRN Fully Convolutional Regression Network
- the Fully Convolutional Regression Network (FCRN) type is disclosed by Xie, W., Noble, J.A., Zisserman, A., Xie, W., Noble, J.A., Microscopy, A.Z., 2016. Computer Methods in Biomechanics and Biomedical Engineering : Imaging & Visualization Microscopy cell counting and detection with fully convolutional regression networks ABSTRACT. Comput. Methods Biomech. Biomed. Eng. IMaging Vis. 1163. doi:10.1080/21681163.2016.1149104 [00234]
- the modified FCRN type is based on the FCRN type, with modifications.
- the CNNs have the following properties:
- FCRN network (by Xie et al) was modified by the following:
- a layer is selected from that a given number of neuron nodes are excluded at random from further calculation.
- the number of such excluded neurons is pre-defined, for example by a percentage (i.e., a dropout parameter).
- GAP global average pooling
- FC fully connected
- GSP global sum pooling
- (i) Class parameters Depending on the use case (cf. FIG. 4), the input and output is differentiated into insect classes. As the classes are known in advance (i.e. prior to operating the CNN), the CNN learns different parameters for different classes (in the training phase) and applies different parameters for different classes (in the production phase). Details are explained in connection with FIGS. 11-12.
- Function parameters indicate the type of operation that the CNN has to apply.
- function parameters can trigger the CNN to estimate density maps (e.g., sum of pixel values indicate the number of objects/insects), to perform particular pre-processing (e.g., to apply Gaussian around a centroid to obtain a kernel) and others.
- the input size parameter i.e., input dimension
- the input dimension defines the tile dimension of a tile that is being processed.
- the input dimension is a tile (or "patch") of 256 x 256 pixels (cf. the discussion regarding tiles, in FIG. 9).
- Activation parameters indicate the type of activation functions, such as sigmoid or softmax (i.e., normalized exponential function) or others.
- Loss function parameters are used to optimize the CNN for accuracy.
- the loss function parameter indicates the difference between the ground truth and the estimation (e.g., number of insects on a leaf manually counted through annotations vs the number of insects estimated by the CNN).
- Loss functions can be defined, for example, by mean functions.
- Auxiliary parameters can be used to deal with technical limitations of the computers.
- computer 201/202 that implements the CNNs may use floating point numbers, with a maximum highest number of 65536.
- CNN 271/272 It may be problematic that CNN 271/272 is not capable of learning what information has to be learned. This is because the contrast (in a density map) between insect (pixel activation of 0.0067) and "no insect" (pixel activation of 0.00) is relatively small. Applying a scale factor increases the contrast in the density maps, and eases the density map estimations (i.e., with integrals over images indicating the number of objects).
- the scale factor can be introduced as auxiliary parameter. For example, all pixel values may be multiplied by the factor 50.000 at the input, and all output values (i.e., insect counts) would be divided by that factor at the output. The factor just shifts the numerical value into a range in that the computer operates more accurately. The mentioned factor is given by way of example, the person of skill in the art can use a different one.
- CNN 261/262 (to detect leaves) is a CNN of the DenseNet type.
- the loss function can be a "binary_crossentropy" function.
- the activation of the last layer can use a "softmax” function.
- the tile dimensions i.e. the dimensions of the input and output image
- CNN 271/272 (i.e. the CNN to detect insects) is a CNN of the
- FCRN type FCRN type.
- the loss functions can be defined, for example, by means functions, such as Mean Absolute Error, or Mean Square Error.
- the tile dimensions can be 256 x 256 pixels.
- FIGS. 11A and 11B illustrate a pixel value filter that is implemented as part of a layer in CNN 271/272 in implementations that differentiate more than two insect classes (c).
- FIG. 11A focuses on a filter that takes individual pixels into account
- FIG. 11B focuses on a filter that takes sets of adjacent pixels (or tile segments) into account.
- the filter can be applied to properties of individual pixels (in the example: the color, in FIG. 11A), and the filter can also be applied to properties of pixel pluralities (in the example: a texture made my multiple pixels, in the pixel-group filter of FIG. 11B).
- FIG. 11A illustrates tile 401-k/402-k as input tile, and on the right side, FIG. 11A illustrates 401-k/402-k as output tiles, differentiated for class (first) and for class (second).
- the filter condition can be implemented, for example, such that pixels from the input are forwarded to the output if the pixel values comply with color parameters Red R(c), Green G(c) and Blue B(c).
- the conditions can be AND-related.
- the color parameters are obtained by training (the insect classes annotated, as explained above).
- the insects in the (second) class should be "blue", so that the parameters are R(first) ⁇ 0.5, G(first) > 0.0, and B(first) > 0.5. Such an insect is illustrated at the lower part of the input tile, again here much simplified with 3 pixels.
- the filter is conveniently implemented as a convolutional layer in CNN 271/272 (the filter filtering tiles, cf. FIG. 11B), but the filter can also be implemented before the splitter module (cf. FIG. 9, the filter for pixels, cf. FIG. 11A).
- the filter can be part of the processing channel of CNN 271/272 before the layer(s) that creates the density maps. Therefore, the density maps are class specific.
- the color parameters Red R(c), Green G(c) and Blue B(c) are just examples for parameters that are related to pixels, but the person of skill in the art can use further parameters such as transparency (if coded in images) etc.
- FIG. 11B illustrates a pixel-group filter that takes neighboring (i.e. adjacent pixels) into account.
- the pixel-group filter uses convolution, it is implemented within CNN 271/272 that processes tiles 401-k/401-k (i.e., after splitter module 241/242).
- FIG. 11B repeats the much simplified "3-pixel-insect" from the left edge of FIG. 11A.
- Segment #10 is being convoluted (with a particular convolution variable, e.g., 3 pixels) to modified segment #10'. Thereby, the pixels values (of the 9 pixels) change.
- segment #10 can have the pixel values (1, 0, 1, 0, 1, 0, 0, 0) and segment #10' can have pixel values (0.8, 0.1, 0.0, 0.1, 0.2, 0.7, 0.1, 0.7, 0.0).
- FIG. 11B illustrates the pixels in "black” or "white” only, the pixels with values over 0.5 are illustrated "black”.
- FIG. 11B also illustrates a further implementation detail.
- the modified segments can be encoded in segment-specific values.
- the figure illustrates such values (by arbitrary numbers) "008", “999” and "008" for segments #1', #10' and #16', respectively.
- the filter criteria can then be applied to the segment-specific values.
- CNN 271/272 can then perform subsequent processing steps by using the segment codes (the segment-specific values). This reduces the number of pixels to be processed (simplified, by a factor that corresponds to the number of pixels per segment, with 9 in the illustrative example). In one of the last layers, CNN 271/272 can then apply decoding.
- FIG. 12 illustrates CNN 271/272 with branches for particular classes (1), (2), (3) and (4).
- the insect-annotations can specify the class (species and growing stage). Training the CNN is performed for channels (or branches) separately.
- channels or branches
- FIG. 12 there are 4 channels corresponding to 4 classes.
- FIG. 12 illustrates combined density maps that combiner 262 obtains by combining maps 502-1 to 502-K, separately for each branch to density maps 555(1), 555(2), 555(3) and 555(4).
- K 36 tiles combined (into one map 555), this number is just selected for simplicity.
- Density maps 502 that indicate the presence of an insect (in the particular class) are illustrated with a dot.
- the overall number of insect is NEST
- CNN 261 is enabled to segment leaves (by becoming CNN 271, method 601B) and enabled to count insects (by becoming CNN 272, method 701B).
- CNN 272 provides NEST (the estimated number of insects per leaf for particular plant-image 412) as the output.
- NEST the estimated number of insects per leaf for particular plant-image 412
- the combination of CNN 262 and CNN 272 would calculate NEST to be exactly the so-called ground truth number NGT: here the number of insects sitting on the particular main leaf of the plant (from that farmer 192 has taken image 412).
- the difference between NEST and NGT would indicate how accurate camera 312 and CNNs 262/272 are performing.
- FIG. 13 illustrates a set of insect-annotated images 472 (i.e., resulting from insect- annotations as in FIG. 6), with sub-sets: a sub-set for training CNN 271 (cf. FIG. 1), a sub-set for validating CNN 272 (cf. FIG. 2), and a sub-set for testing CNN 272.
- the sub-sets have cardinalities Si, S2 and S3, respectively.
- FIG. 13 takes the insect-annotated leaf-images 471 as an example only.
- the person of skill in the art can fine- tune training CNN 271 for the leaf segmentation accordingly.
- CNN 261/271 have been trained with the Si images of the training sub-set to become trained-CNN 262, 272 Trained-CNN 262, 272 have been used to estimate NEST for the S2 images of the validation sub-set.
- the S2 values NGT are known from the insect-annotations. If for a particular image, NEST is higher than NGT CNNS 262/272 have counted more insects that present in reality.
- FIG. IS illustrates a simplified graph 503 showing NEST on the ordinate versus ground truth NGT on the abscissa, with a dot identifying an (NEST, NGT) pair.
- Graph 503 is simplified in illustrating 9 dots only (instead of, for example, 54 or 123).
- Most of the dots are located approximately along regression line 504. The
- a metric can be defined as Mean Absolute Error (MAE), or
- NEST and NGT are obtained as the average of the S 3 images.
- a further metric can be defined as Mean Square Error (MSE), or
- MSE ROOT [(NEST - NGT) 2 ]. Again, NEST AND NGT for all S 3 has to be taken into account (i.e. [ ] being the sum of ( ) 2 for all S3).
- Differentiating the main leaf from its adjacent leaves (or neighbor leaves) can be implemented by known methods as well (among them feature extraction). For leaves that are green over a non-green ground, color can be used as a differentiator. However, such an approach would eventually fail for "green” over “green” situations, for example, when one leaf overlaps another leaf.
- counting insects can be implemented by other known approaches, such as by the above-mentioned candidate selection with subsequent classification.
- FIG. 3 Shortly returning to FIG. 1, it illustrates computer 303 as a further computing function.
- FIG. 18 The figures illustrates computers 301 and 303 in parallel, because both receive images and both provide annotated images. While computer 301 interacts with expert user 191, computer 303 performs computer-implemented method (800, cf. FIG. 18) without expert user 191. It is however possible to implement both computing functions 301/303 by a single physical computer.
- computer 303 (method 800) provides images 473 (insect-annotated, but not leaf-annotated).
- the images at the input of computer 303 are leaf-images (i.e., images that show the main leaf as explained above).
- the description explains approaches to obtain such leaf-images, for example, by using leaf segmentation 601A, or 601B (cf. the overview in FIG. 3).
- insect-annotated leaf-images 471 are images with human- made annotations.
- insect-annotated leaf-images 473 are images with computer-made annotations (method 800).
- a first training branch uses images 471 and a second training branch uses images 473.
- a branch under "full supervision” and there is branch under "semi-supervision”.
- the results of both branches can be compared (for example, by investigating the loss function) so that method 800 can be adjusted.
- the training phase (to train CNN 271 to CNN 272) is enhanced by method 800 that provide images 473 (that are used in addition to image 471, the first and second branches).
- Method 800 can be regarded as an auxiliary method.
- computer 301 interacts with expert user 191 to obtain insect- annotations, resulting in image 471.
- computer 303 - without interacting with the expert - provides insect-annotated plant images 473, being computer-annotated images.
- Cameras 311 and 312 are optimized to take images targeted for human viewers (e.g., photos that show people, pet animals, and/or buildings).
- the colors of the photos are optimized for the human eye.
- the images are not optimized to count insects on plants, not by human users, not by computers.
- FIG. 14 illustrates pixel-by-pixel color-coding and illustrates the concept of assigning pixel colors to binary pixel values. Much simplified, there should be a leaf (here symbolized by a large square) with an insect (symbolized by a circle in the center of the square).
- leaf pixels 477 for the leaf
- insect pixels 478 for the insect
- the quotations merely point to a simplification.
- the leaf pixels 477 and the insect pixels 478 are coded in a particular color-coding.
- the coding uses 3 components (i.e., a "color space").
- the example uses RGB-coding in that the color components R, G and B are coded by real numbers (e.g., in closed intervals [0,1]).
- the (R, G, B) color space would be reduced to a (R, G) planar color space 479, with component values (0, 1) for pixels 477, and (1, 0) for pixels 478.
- the description uses the term "space” in the mathematical sense. For 3 color components, the space would be a 3D space. Removing a component changes the 3D space to a 2D "plane" (i.e. flat surface in mathematical terms). Since the term "plane" is frequently used to describe technologies such image processing or printing, the description herein uses the term "planar color space”.
- the color of the pixels would be on opposite ends of planar color space 479, with an easy to differentiate contrast 485.
- FIG. 14 illustrates binary values as components in “white” and “black”, but there is no need to show images to a user.
- Binary values differentiate the leaf from the insect.
- insects are not “red” and they are not “white” (even if the name may suggest that). As insects tend to camouflage, they may be in colors that are similar to that of the leaf. It is noted that camouflage would be sensitive to the eyes of insect-eating birds (or the like) but not sensitive to human eyes and not sensitive to cameras 311/312 (cf. FIGS. 1-2). [00325] The description now explains how a color-code change addresses this problem. The sensitivity (of the computer) to small color differences that are typical for leaf/insect images is thereby increased. Color space transformation
- FIG. 15 illustrates pixel-by-pixel color-coding and assigning colors to binary pixel values, with changing the color-coding involving color space transformation and color clustering. The figure takes over the leaf/insect example of FIG. 14.
- This coding is first color-coding 481 (i.e. a coding in a color space), here again RGB- coding.
- the change comprises a space-transformation, and a subsequent space- to-plane reduction (i.e., 3D space to 2D planar space)
- XYZ is a color space defined by CIE (Commission Internationale de I'Eclairage, International Commission on Illumination) in standards (ISO/CIE 11664-4:2019 Colorimetry). (The letters X, Y, Z are not identical with location coordinates).
- the subsequent space-to-planar-space reduction is simply performed by disregarding the Z-values.
- the Z-values for the leaf and for the insects are substantially equal.
- the Z-values do not substantially contribute to differentiating (leaf/insect).
- Such a reduction from space (X, Y, Z) to planar space (X, Y) is similar to the space/planar-space reduction of FIG. 14.
- the color-values (X, Y) can be illustrated in planar color space 479, and the line between both can be taken as threshold 485.
- the (X, Y) color-coding is second color-coding 482.
- threshold 485. For example, pixels with Y > 0.5 would be coded to (0, 1) and pixels with Y ⁇ 0.5 would be coded to (1, 0).
- the pixel(s) above (or at) the threshold belong to a first color cluster, and the pixel(s) below the a second color cluster. Color clusters are explained with more detail in FIG. 16.
- the coding change (transformation, reduction) enhances the contrast between the leaf and the insect parts and makes it easier to assign pixels to binary values.
- the leaf-insect contrast is higher (than in the first color-coding). In other words, the clusters are differentiated by color contrast.
- FIG. 15 explains a simplified example with pixels in two colors, it is noted that images may have pixels in a couple of hundred component combinations. A human user may easily differentiate the colors into “green” or “white” colors, but computer 303 takes a different approach.
- FIG. 16 illustrates clustering as a tool for assigning colors to binary pixel values.
- the variety of similar colors is reflected by different color components.
- the pixel colors After transformation (and reduction to a single planar space, second color-coding 482), the pixel colors would be distributed in clusters. For example, a first cluster would be more with "green” pixels, a second cluster would be more with the "white” pixels.
- the "colors" are just illustrative examples.
- FIG. 16 therefore illustrates planar color space 479 (X, Y) with occurrences of pixels with different colors.
- the small squares in planar color space 479 indicate particular pixel colors (i.e. particular second color-coding).
- Different color groups can be differentiated by color clustering.
- Clustering techniques are available to the skilled person.
- the skilled person can use support vector machines (SVM), Bayesian classifiers, Kmeans, k-nearest neighbors (KNN) in whatever dimensionality.
- SVM support vector machines
- Bayesian classifiers Kmeans
- KNN k-nearest neighbors
- Line 485 between both clusters corresponds to the threshold.
- the line is illustrated as a free form line to illustrates that the "threshold" is not necessarily a particular value.
- a (N-l)D structure In ND space, a (N-l)D structure would determine the separation between clusters. For example, in a 3D space, planes would determine the separation.
- the binary classification depends on the cluster (leaf color cluster vs. insect color cluster). It is noted that the semantic (e.g., leaf or insect) does not matter.
- the computer that perform clustering e.g., computer 303 does not have to know the semantics, for the computer the clusters do not stand for leaf or insect, they are just color codes.
- Variations of the clustering approach are possible. A first approach disregards outliers can be disregarded (e.g., relatively few pixels in a color 486), a second approach disregards color that belong to other areas. For example, a some areas on the leaf would be spots in "brown" color (semantics: the insects have eaten holes into the leaf so that the background becomes visible). It is relatively easy to introduce minimal expert interaction: the human expert would identify these third color spots to be mapped to either the first or to the second binary values. In other words, few expert supervision is helpful when 3 or more cluster have to be mapped to binary values.
- FIG. 17 illustrates a sequence of images, starting from image 413-A being a leaf- image, and ending at insect-annotated leaf image 473. By in large, the sequence follows the step of method 800.
- image 413-A is a leaf-image that can be obtained by performing method 602B. It does not matter, if image 413-A shows a single leaf only or if the area surrounding that single leaf are identified as to be ignored.
- Image 413-A shows leaf 121 (cf. FIG. 1) by symbolizing the leaf border (or leaf edge, or leaf margin) by a plain line. For convenience, this plain line is kept for the other images of the sequence. The pixels outside the leaf (border) are being ignored.
- Lines 121' illustrate folds and nerves of the leaf.
- insects 131 shown on leaf 121 (or rather things that look like insects).
- the figure shows a plurality of 8 insect 131, and this low number is just used to keep the figure simple. These insects are not to be counted, but to be identified and located.
- the insects are illustrated here by circle symbols, just for simplicity, oval-shaped insects are introduced in FIGS. 1-2. The reader can imagine to have "white” spots on "green” leaves.
- FIG. 1 For computer 303 (FIG. 1, that executes the method) the semantical difference between "leaf" and "insect” does not (yet) matter.
- a small hexagon symbolizes a further object 161 that is located on leaf 121.
- the object could be a water drop or a dust particle.
- Image 413-A is color-coded in the first color-coding, for example in RGB because camera 311 (cf. FIG. 1) uses that coding.
- the figure also uses image 413-A to illustrate pixel coordinates (i, j). The i, j notation distinguishes pixel coordinates (of the image) from color-coordinate (of the color space, cf. FIGS. 14-16).
- Images 413-B and 413-C are images in different color-coding after transformation, as used herein RGB-to-XYZ transformation.
- Image 413-B was obtained by a RGB-to-XYZ transformation (of image 413-A). Just transforming the coding does change the appearance of the image, and the human user would perceive colors differently.
- Image 413-C is obtained from image 413-B by modifying the XYZ-coding. Modifying means to disregard one channel (X, Y, or Z).
- Image 413-C is in the second color-coding
- Image 413-C is in the second color-coding, for which the contrast is higher than for the first coding.
- the figures illustrate the contrast enhancement by showing the insects with bold circles. In the second color-coding, the contrast has been enhanced by the reduction from (X, Y, Z) to (X, Y) as explained. In other words, channel Z was disregarded.
- image 413-C does not show object 161 because in the second color-coding for object 161 is similar to that of leaf 121. It is advantageous that the coding change filters out such objects early during processing.
- Image 413-D is in binary-coding. This is just illustrated by the leaves in “green” coded to (0,1) and the insects in “white” coded to (1, 0). The figure presents the color in negative, just to keep references visible. [00363] (1,0) coded pixels can be grouped together to form areas of contiguous pixel
- the areas have borders where the pixel coding changes from (1, 0) to (0, 1).
- the figure illustrates areas 1 to 8.
- the areas can be quantified by numeric values (AREA-SIZE), for example, by counting the number of pixels (per area), by counting the highest number of pixels in one pixel coordinate (i, or j) etc. Other features of the regions that can be used as well, among them centroid, convexity, circularity, inner bounding box, outer bounding box, Euler number, etc. [00366] According to a size criterion (numerical value(s)), the areas are classified into insect areas and non-insect areas.
- an area is an insect area for MIN ⁇ AREA-SIZE ⁇ MAX. Areas 1 and 5 are too small or too large, but areas 2-4 and 6-8 are insect areas.
- the size criterion can be obtained by the computer, without the need to interact with human supervisors.
- the criterion can also be obtained from an expert user (cf. FIG. 1, 191), even by re using results of the above-describe annotation process (cf. 701A in FIG. 3).
- the expert annotations provide size information as a side-product.
- the non-insect areas are disregarded (or filtered out).
- the insect areas have center pixels (e.g., the pixel in that the longest line with constant i crosses the longest line with constant j). In other words, there is a collection of center pixels with center pixel coordinates. (In the example there are 6 center pixels, for 6 insect areas).
- the pixels surrounding the center pixels show insect with some leaf parts.
- Computer 303 now takes over the surrounding pixel from the original image 413-A (in first color-coding with all 3 color channels) (or from image 413-B, with XYZ)
- the pixels in a square around that center are forming "tiles".
- the size of the tiles is standardized, for example 96 pixel in direction of coordinate i and 96 pixel in direction of coordinate j.
- a tile can have 96 x 96 pixels.
- the pixel number is just taken for convenience because experiments showed that the size is big enough to contain an insect completely.
- variability of the size exists due to different resolution images and acquisition conditions.
- FIG. 17 focuses on insect areas to obtain tiles (with insects), the person of skill in the art is able to obtain non-insect tiles accordingly so that training can use such non insect tiles as well.
- FIG. 4 differentiates whitefly insects into “insect (1)” and “no-insect”, and in a higher granularity into classes (1) to (4).
- computer 301 (cf. FIG. 1, the co-operating computing function) has images for that annotations do differentiate already, for example, "alive whitefly", “dead whitefly” and "no whitefly”.
- FIG. 17 illustrates this optional result- sharpening approach:
- the person of skill in the art can decide in advance who to classify tiles such as tile 2.
- FIG. 18 illustrates a method-flow chart of computer-implemented method 800 for generating a training set with annotated images.
- annotated images 473 are to be used to train a convolutional neural network (CNN) for quantifying plant infestation by estimating the number NEST of insects 132 on leaves 122 of plants 112.
- Method 800 is being performed by computer 303 (cf. FIG. 1) and is explained be referring to FIGS. 14-17.
- Computer 303 receives 810 leaf-images 413-A showing leaves 121 and showing insects 131 on the leaves 121.
- Leaf-images 413-A are coded in first color-coding 481.
- Computer 303 changes 820 the color-coding of pixels 477, 478 of the leaf-images to second color-coding 482. Contrast 485 between the pixels for the insects and the pixels for the leaves is higher in the second color-coding than in the first color-coding.
- Computer 303 assigns 830 pixels 477/478 in second color-coding 482 to a first binary value or to a second binary value. Color clusters are related to the first and second binary values: the pixels associated with the second color cluster are assigned to the second binary value. [00383] Computer 303 differentiates 840 - in the leaf-images with binary coding 413-D - areas with contiguous pixels in the first binary value into into non-insect areas 1, 5 and insect areas 2, 3, 4, 6, 7, 8 by an area size criterion.
- Computer 303 identifies 850 pixel-coordinates of insect areas 2, 3, 4, 6, 7, 8 wherein the pixel-coordinates identify rectangular tile-areas with insects in the center [00385] Computer 303 annotates 860 leaf-images 414-A (in first color-coding 481) by assigning the pixel-coordinates (i, j) to corresponding tile-areas to obtain the annotated image 473.
- FIG. 8 illustrates an image being split into tiles or sub-regions.
- the tiles are already in a size that is suitable as input to CNN 271.
- the tiles correspond to tiles 401-k (alpha, with insects) and 401-k (gamma, without insects).
- FIG. 19 illustrates CNN 271 being trained.
- CNN 271 is trained by images 471 (cf. FIG. 1) and CNN 271 turns into CNN 272.
- CNN 271 can be trained by images 473 (insect-by-computer-annotated images) but the contribution of theses images 473 (to the training) is taken into account with less emphasis.
- training a CNN involves the calculation of loss values.
- the description has mentioned to train CNN 271 by using a loss-function that is the mean absolute error (MAE) or that is the mean square error (MSE).
- MAE mean absolute error
- MSE mean square error
- FIG. 19 illustrates this approach by switch symbols at the input of the CNN and a the part of the CNN that calculates the loss function. Both LOSS_l and LOSS_2 are calculated according to the same formulas, but both values are kept separately.
- the lambda value is smaller than 1 so that the contribution of the computer- annotated images is less trusted.
- FIG. 19 illustrates a first sub-set ⁇ 471 ⁇ and a second sub-set ⁇ 473 ⁇ , and illustrates that the images go into the network consecutively.
- the network calculates the loss function (described above) and uses the loss function (in a feedback loop, with the overall LOSS) to fine-tune the parameter (or layer-specific values), as explained above.
- image 412 could be taken otherwise, for example by aircraft flying over the field.
- UAV unmanned aerial vehicle
- the tiles have smaller dimensions than the images, and the tile dimensions correspond to the input layer dimension of CNN 271/272.
- the biological objects e.g., the insects
- the physical sizes of the biological object are limited to a maximal physical size. In the extreme case (maximum), the biological object of the largest allowable physical size would correspond to the representation of that object on a single tile.
- the relation of the physical size of the biological objects (132) to the physical size of the parts (122) is such that the representation of the biological objects on the part-images (422) are such that the representation is smaller than the tile dimension.
- the image resolution i.e., the number of pixels per physical dimension.
- the minimum In the extreme case (minimum), the biological object of the smallest allowable size would be represented (in theory) by one pixel. More practical sizes have been explained above (cf. FIG. 5, 30 pixels times 20 pixels). This translates to absolute minimum size of biological objects that can be recognized.
- the biological objects have at least 0.1 mm in diameter, preferably at least 0.5 mm, more preferably at least 1 mm, most preferably at least 5 mm in diameter.
- the biological objects (132) are (or were) living organisms that are located on the parts (122) (of the plant), or the biological objects are traces by that organisms. (Optionally, the organism may be considered as no longer living, cf. the example with the pupa). More in detail, the biological objects (132) on the parts (122) (of the plant) are selected from the following: insects, arachnids, and mollusca. In an alternative, the biological objects are selected from spots or stripes on the surface of the plant parts. In that alternative, it does not matter if the objects are considered to be organisms or not, spots or stripes can be disease symptoms. For example, brown spots or brown stripes on a plant part indicates that the plant is potentially damaged.
- FIG. 20 illustrates an example of a generic computer device 900 and a generic mobile computer device 950, which may be used with the techniques described here
- Computing device 900 is intended to represent various forms of digital computers, such as laptops, desktops, workstations, personal digital assistants, servers, blade servers, mainframes, and other appropriate computers.
- Generic computer device may 900 correspond to computers 201/202 of FIGS. 1-2.
- Computing device 950 is intended to represent various forms of mobile devices, such as personal digital assistants, cellular telephones, smart phones, and other similar computing devices.
- computing device 950 may include the data storage components and/or processing components of devices as shown in FIG. 1.
- Computing device 900 includes a processor 902, memory 904, a storage device
- the processor 902 can process instructions for execution within the computing device 900, including instructions stored in the memory 904 or on the storage device 906 to display graphical information for a GUI on an external input/output device, such as display 916 coupled to high speed interface 908.
- multiple processors and/or multiple buses may be used, as appropriate, along with multiple memories and types of memory.
- multiple computing devices 900 may be connected, with each device providing portions of the necessary operations (e.g., as a server bank, a group of blade servers, or a multi-processor system).
- the memory 904 stores information within the computing device 900.
- the memory 904 is a volatile memory unit or units.
- the memory 904 is a non-volatile memory unit or units.
- the memory 904 may also be another form of computer-readable medium, such as a magnetic or optical disk.
- the storage device 906 is capable of providing mass storage for the computing device 900.
- the storage device 906 may be or contain a computer- readable medium, such as a floppy disk device, a hard disk device, an optical disk device, or a tape device, a flash memory or other similar solid state memory device, or an array of devices, including devices in a storage area network or other configurations.
- a computer program product can be tangibly embodied in an information carrier.
- the computer program product may also contain instructions that, when executed, perform one or more methods, such as those described above.
- the information carrier is a computer- or machine-readable medium, such as the memory 904, the storage device 906, or memory on processor 902.
- the high speed controller 908 manages bandwidth-intensive operations for the computing device 900, while the low speed controller 912 manages lower bandwidth-intensive operations.
- the high speed controller 908 is coupled to memory 904, display 916 (e.g., through a graphics processor or accelerator), and to high-speed expansion ports 910, which may accept various expansion cards (not shown).
- low-speed controller 912 is coupled to storage device 906 and low-speed expansion port 914.
- the low-speed expansion port which may include various communication ports (e.g., USB, Bluetooth, Ethernet, wireless Ethernet) may be coupled to one or more input/output devices, such as a keyboard, a pointing device, a scanner, or a networking device such as a switch or router, e.g., through a network adapter.
- the computing device 900 may be implemented in a number of different forms, as shown in the figure. For example, it may be implemented as a standard server 920, or multiple times in a group of such servers. It may also be implemented as part of a rack server system 924. In addition, it may be implemented in a personal computer such as a laptop computer 922.
- components from computing device 900 may be combined with other components in a mobile device (not shown), such as device 950.
- a mobile device not shown
- Each of such devices may contain one or more of computing device 900, 950, and an entire system may be made up of multiple computing devices 900, 950 communicating with each other.
- Computing device 950 includes a processor 952, memory 964, an input/output device such as a display 954, a communication interface 966, and a transceiver 968, among other components.
- the device 950 may also be provided with a storage device, such as a microdrive or other device, to provide additional storage.
- a storage device such as a microdrive or other device, to provide additional storage.
- Each of the components 950, 952, 964, 954, 966, and 968 are interconnected using various buses, and several of the components may be mounted on a common motherboard or in other manners as appropriate.
- the processor 952 can execute instructions within the computing device 950, including instructions stored in the memory 964.
- the processor may be implemented as a chipset of chips that include separate and multiple analog and digital processors.
- the processor may provide, for example, for coordination of the other components of the device 950, such as control of user interfaces, applications run by device 950, and wireless communication by device 950.
- Processor 952 may communicate with a user through control interface 958 and display interface 956 coupled to a display 954.
- the display 954 may be, for example, a TFT LCD (Thin-Film-Transistor Liquid Crystal Display) or an OLED (Organic Light Emitting Diode) display, or other appropriate display technology.
- the display interface 956 may comprise appropriate circuitry for driving the display 954 to present graphical and other information to a user.
- the control interface 958 may receive commands from a user and convert them for submission to the processor 952.
- an external interface 962 may be provide in communication with processor 952, so as to enable near area communication of device 950 with other devices.
- External interface 962 may provide, for example, for wired communication in some implementations, or for wireless communication in other implementations, and multiple interfaces may also be used.
- the memory 964 stores information within the computing device 950.
- the memory 964 can be implemented as one or more of a computer-readable medium or media, a volatile memory unit or units, or a non-volatile memory unit or units.
- Expansion memory 984 may also be provided and connected to device 950 through expansion interface 982, which may include, for example, a SIMM (Single In Line Memory Module) card interface. Such expansion memory 984 may provide extra storage space for device 950, or may also store applications or other information for device 950.
- SIMM Single In Line Memory Module
- expansion memory 984 may include instructions to carry out or supplement the processes described above, and may include secure information also.
- expansion memory 984 may act as a security module for device 950, and may be programmed with instructions that permit secure use of device 950.
- secure applications may be provided via the SIMM cards, along with additional information, such as placing the identifying information on the SIMM card in a non-hackable manner.
- the memory may include, for example, flash memory and/or NVRAM memory, as discussed below.
- a computer program product is tangibly embodied in an information carrier.
- the computer program product contains instructions that, when executed, perform one or more methods, such as those described above.
- the information carrier is a computer- or machine-readable medium, such as the memory 964, expansion memory 984, or memory on processor 952, that may be received, for example, over transceiver 968 or external interface 962.
- Device 950 may communicate wirelessly through communication interface 966, which may include digital signal processing circuitry where necessary. Communication interface 966 may provide for communications under various modes or protocols, such as GSM voice calls, SMS, EMS, or MMS messaging, CDMA, TDMA, PDC, WCDMA, CDMA2000, or GPRS, among others. Such communication may occur, for example, through radio-frequency transceiver 968. In addition, short-range communication may occur, such as using a Bluetooth, WiFi, or other such transceiver (not shown). In addition, GPS (Global Positioning System) receiver module 980 may provide additional navigation- and location-related wireless data to device 950, which may be used as appropriate by applications running on device 950.
- GPS Global Positioning System
- Device 950 may also communicate audibly using audio codec 960, which may receive spoken information from a user and convert it to usable digital information. Audio codec 960 may likewise generate audible sound for a user, such as through a speaker, e.g., in a handset of device 950. Such sound may include sound from voice telephone calls, may include recorded sound (e.g., voice messages, music files, etc.) and may also include sound generated by applications operating on device 950.
- Audio codec 960 may receive spoken information from a user and convert it to usable digital information. Audio codec 960 may likewise generate audible sound for a user, such as through a speaker, e.g., in a handset of device 950. Such sound may include sound from voice telephone calls, may include recorded sound (e.g., voice messages, music files, etc.) and may also include sound generated by applications operating on device 950.
- the computing device 950 may be implemented in a number of different forms, as shown in the figure. For example, it may be implemented as a cellular telephone 980. It may also be implemented as part of a smart phone 982, personal digital assistant, or other similar mobile device.
- Various implementations of the systems and techniques described here can be realized in digital electronic circuitry, integrated circuitry, specially designed ASICs (application specific integrated circuits), computer hardware, firmware, software, and/or combinations thereof.
- ASICs application specific integrated circuits
- These various implementations can include implementation in one or more computer programs that are executable and/or interpretable on a programmable system including at least one programmable processor, which may be special or general purpose, coupled to receive data and instructions from, and to transmit data and instructions to, a storage system, at least one input device, and at least one output device.
- the systems and techniques described here can be implemented on a computer having a display device (e.g., a CRT (cathode ray tube) or LCD (liquid crystal display) monitor) for displaying information to the user and a keyboard and a pointing device (e.g., a mouse or a trackball) by which the user can provide input to the computer.
- a display device e.g., a CRT (cathode ray tube) or LCD (liquid crystal display) monitor
- a keyboard and a pointing device e.g., a mouse or a trackball
- Other kinds of devices can be used to provide for interaction with a user as well; for example, feedback provided to the user can be any form of sensory feedback (e.g., visual feedback, auditory feedback, or tactile feedback); and input from the user can be received in any form, including acoustic, speech, or tactile input.
- the systems and techniques described here can be implemented in a computing device that includes a back end component (e.g., as a data server), or that includes a middleware component (e.g., an application server), or that includes a front end component (e.g., a client computer having a graphical user interface or a Web browser through which a user can interact with an implementation of the systems and techniques described here), or any combination of such back end, middleware, or front end components.
- the components of the system can be interconnected by any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include a local area network ("LAN”), a wide area network (“WAN”), and the Internet.
- the computing device can include clients and servers. A client and server are generally remote from each other and typically interact through a communication network.
- client and server arises by virtue of computer programs running on the respective computers and having a client-server relationship to each other.
Landscapes
- Engineering & Computer Science (AREA)
- Theoretical Computer Science (AREA)
- Physics & Mathematics (AREA)
- General Physics & Mathematics (AREA)
- Evolutionary Computation (AREA)
- General Health & Medical Sciences (AREA)
- Health & Medical Sciences (AREA)
- Artificial Intelligence (AREA)
- Software Systems (AREA)
- Computing Systems (AREA)
- Multimedia (AREA)
- Life Sciences & Earth Sciences (AREA)
- Computer Vision & Pattern Recognition (AREA)
- Biomedical Technology (AREA)
- Molecular Biology (AREA)
- Data Mining & Analysis (AREA)
- General Engineering & Computer Science (AREA)
- Biophysics (AREA)
- Computational Linguistics (AREA)
- Mathematical Physics (AREA)
- Medical Informatics (AREA)
- Databases & Information Systems (AREA)
- Biodiversity & Conservation Biology (AREA)
- Nuclear Medicine, Radiotherapy & Molecular Imaging (AREA)
- Quality & Reliability (AREA)
- Radiology & Medical Imaging (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Bioinformatics & Computational Biology (AREA)
- Evolutionary Biology (AREA)
- Image Analysis (AREA)
- Catching Or Destruction (AREA)
- Image Processing (AREA)
Applications Claiming Priority (3)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| EP19200657.5A EP3798899A1 (de) | 2019-09-30 | 2019-09-30 | Quantifizierung von pflanzenbefall durch schätzung der anzahl der insekten auf blättern durch neuronale faltungsnetze, die dichtekarten bereitstellen |
| EP20158881.1A EP3798901A1 (de) | 2019-09-30 | 2020-02-21 | Quantifizierung des pflanzenbefalls durch schätzung der zahl der insekten auf blättern, durch neuronale faltungsnetzwerke, die trainingsbilder verwenden, die von einem halbüberwachten ansatz aufgenommen werden |
| PCT/EP2021/054233 WO2021165512A2 (en) | 2019-09-30 | 2021-02-19 | Quantifying plant infestation by estimating the number of biological objects on leaves, by convolutional neural networks that use training images obtained by a semi-supervised approach |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| EP4107656A2 true EP4107656A2 (de) | 2022-12-28 |
| EP4107656B1 EP4107656B1 (de) | 2025-04-09 |
Family
ID=68109134
Family Applications (4)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP19200657.5A Withdrawn EP3798899A1 (de) | 2019-09-30 | 2019-09-30 | Quantifizierung von pflanzenbefall durch schätzung der anzahl der insekten auf blättern durch neuronale faltungsnetze, die dichtekarten bereitstellen |
| EP20158881.1A Ceased EP3798901A1 (de) | 2019-09-30 | 2020-02-21 | Quantifizierung des pflanzenbefalls durch schätzung der zahl der insekten auf blättern, durch neuronale faltungsnetzwerke, die trainingsbilder verwenden, die von einem halbüberwachten ansatz aufgenommen werden |
| EP20781367.6A Active EP4038542B1 (de) | 2019-09-30 | 2020-09-29 | Quantifizierung von pflanzenbefall durch schätzung der anzahl der insekten auf blättern durch neuronale faltungsnetze, die dichtekarten bereitstellen |
| EP21705986.4A Active EP4107656B1 (de) | 2019-09-30 | 2021-02-19 | Quantifizierung des pflanzenbefalls durch schätzung der zahl der insekten auf blättern, durch neuronale faltungsnetzwerke, die trainingsbilder verwenden, die von einem halbüberwachten ansatz aufgenommen werden |
Family Applications Before (3)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP19200657.5A Withdrawn EP3798899A1 (de) | 2019-09-30 | 2019-09-30 | Quantifizierung von pflanzenbefall durch schätzung der anzahl der insekten auf blättern durch neuronale faltungsnetze, die dichtekarten bereitstellen |
| EP20158881.1A Ceased EP3798901A1 (de) | 2019-09-30 | 2020-02-21 | Quantifizierung des pflanzenbefalls durch schätzung der zahl der insekten auf blättern, durch neuronale faltungsnetzwerke, die trainingsbilder verwenden, die von einem halbüberwachten ansatz aufgenommen werden |
| EP20781367.6A Active EP4038542B1 (de) | 2019-09-30 | 2020-09-29 | Quantifizierung von pflanzenbefall durch schätzung der anzahl der insekten auf blättern durch neuronale faltungsnetze, die dichtekarten bereitstellen |
Country Status (8)
| Country | Link |
|---|---|
| US (2) | US12073327B2 (de) |
| EP (4) | EP3798899A1 (de) |
| CN (2) | CN114467122A (de) |
| AR (1) | AR120119A1 (de) |
| BR (2) | BR112022002456A2 (de) |
| CA (2) | CA3150808A1 (de) |
| ES (2) | ES3010674T3 (de) |
| WO (2) | WO2021063929A1 (de) |
Families Citing this family (22)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP3798899A1 (de) * | 2019-09-30 | 2021-03-31 | Basf Se | Quantifizierung von pflanzenbefall durch schätzung der anzahl der insekten auf blättern durch neuronale faltungsnetze, die dichtekarten bereitstellen |
| TWI798655B (zh) * | 2020-03-09 | 2023-04-11 | 美商奈米創尼克影像公司 | 缺陷偵測系統 |
| CA3171844A1 (en) * | 2020-03-16 | 2021-09-23 | Basf Se | Quantifying biotic damage on plant leaves, by convolutional neural networks |
| US12423659B2 (en) * | 2020-05-05 | 2025-09-23 | Planttagg, Inc. | System and method for horticulture viability prediction and display |
| US11748984B2 (en) * | 2020-05-05 | 2023-09-05 | Planttagg, Inc. | System and method for horticulture viability prediction and display |
| CN113096100B (zh) * | 2021-04-15 | 2023-08-22 | 杭州睿胜软件有限公司 | 用于植物病症诊断的方法和植物病症诊断系统 |
| CN113221740B (zh) * | 2021-05-12 | 2023-03-24 | 浙江大学 | 一种农田边界识别方法及系统 |
| CN113256578B (zh) * | 2021-05-18 | 2025-01-21 | 河北农业大学 | 一种入侵植物危害检测方法 |
| US11838067B2 (en) * | 2021-12-23 | 2023-12-05 | Dish Network L.L.C. | Signal interference prediction systems and methods |
| WO2023170975A1 (ja) * | 2022-03-11 | 2023-09-14 | オムロン株式会社 | 学習方法、葉状態識別装置、およびプログラム |
| CN116958013B (zh) * | 2022-04-12 | 2025-09-16 | 腾讯科技(深圳)有限公司 | 图像中对象数量的估计方法、装置、介质、设备及产品 |
| CN115035131B (zh) * | 2022-04-24 | 2025-08-22 | 南京农业大学 | U型自适应est的无人机遥感图像分割方法及系统 |
| CN115126686B (zh) * | 2022-08-31 | 2022-11-11 | 山东中聚电器有限公司 | 用于植保无人机载隔膜泵控制系统 |
| CN116934143A (zh) * | 2023-07-14 | 2023-10-24 | 杭州睿胜软件有限公司 | 植物的健康状态评估方法、装置及计算机可读存储介质 |
| US20250102360A1 (en) * | 2023-09-26 | 2025-03-27 | Arizona Board Of Regents On Behalf Of Arizona State University | Terrestrial observing network for digital twins: real-time 3d mapping of metric, semantic, topological, and physicochemical properties for optimal environmental monitoring |
| CN117372881B (zh) * | 2023-12-08 | 2024-04-05 | 中国农业科学院烟草研究所(中国烟草总公司青州烟草研究所) | 一种烟叶病虫害智能识别方法、介质及系统 |
| CN118037821B (zh) * | 2024-03-14 | 2024-12-13 | 湖北麦麦农业科技有限公司 | 一种生菜叶片面积检测方法及系统 |
| CN118537608B (zh) * | 2024-06-06 | 2025-02-28 | 广东省农业科学院植物保护研究所 | 基于深度学习的荔枝蒂蛀虫为害叶片检测方法及系统 |
| CN118840738B (zh) * | 2024-07-31 | 2025-03-25 | 博罗罗浮山润心食品有限公司 | 一种豆制品生产过程智能管理方法及其系统 |
| WO2026046813A1 (en) | 2024-08-26 | 2026-03-05 | Bayer Aktiengesellschaft | Locating and classifying arthropods in images |
| EP4703918A1 (de) * | 2024-08-26 | 2026-03-04 | Bayer Aktiengesellschaft | Lokalisierung und klassifizierung von arthropoden in bildern |
| CN118781218B (zh) * | 2024-09-13 | 2024-12-20 | 安徽医科大学第一附属医院 | 基于半监督学习的胃肠影像重建方法、装置及电子设备 |
Family Cites Families (20)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US7496228B2 (en) * | 2003-06-13 | 2009-02-24 | Landwehr Val R | Method and system for detecting and classifying objects in images, such as insects and other arthropods |
| KR101109337B1 (ko) * | 2009-03-24 | 2012-01-31 | 부산대학교 산학협력단 | 자동 해충 인지 및 방제 시스템 및 방법 |
| CN104598908B (zh) * | 2014-09-26 | 2017-11-28 | 浙江理工大学 | 一种农作物叶部病害识别方法 |
| US10349584B2 (en) * | 2014-11-24 | 2019-07-16 | Prospera Technologies, Ltd. | System and method for plant monitoring |
| CN104992439B (zh) * | 2015-06-26 | 2020-02-11 | 广州铁路职业技术学院 | 农作物叶子虫害检测装置及其检测方法 |
| JP6921095B2 (ja) * | 2015-11-08 | 2021-08-25 | エーグロウイング リミテッド | 航空画像を収集及び分析するための方法 |
| CN107292314A (zh) * | 2016-03-30 | 2017-10-24 | 浙江工商大学 | 一种基于cnn的鳞翅目昆虫种类自动鉴别方法 |
| EP3287007A1 (de) * | 2016-08-24 | 2018-02-28 | Bayer CropScience AG | Bekämpfung von schadorganismen auf basis der vorhersage von befallsrisiken |
| CN106504258B (zh) * | 2016-08-31 | 2019-04-02 | 北京农业信息技术研究中心 | 一种植物叶片图像提取方法及装置 |
| US10339380B2 (en) * | 2016-09-21 | 2019-07-02 | Iunu, Inc. | Hi-fidelity computer object recognition based horticultural feedback loop |
| ES2956102T3 (es) * | 2016-10-28 | 2023-12-13 | Verily Life Sciences Llc | Modelos predictivos para clasificar visualmente insectos |
| US10796161B2 (en) | 2017-07-14 | 2020-10-06 | Illumitex, Inc. | System and method for identifying a number of insects in a horticultural area |
| US11263707B2 (en) * | 2017-08-08 | 2022-03-01 | Indigo Ag, Inc. | Machine learning in agricultural planting, growing, and harvesting contexts |
| US10455826B2 (en) * | 2018-02-05 | 2019-10-29 | FarmWise Labs, Inc. | Method for autonomously weeding crops in an agricultural field |
| CN108960310A (zh) * | 2018-06-25 | 2018-12-07 | 北京普惠三农科技有限公司 | 一种基于人工智能的农业病虫害识别方法 |
| CN108921849A (zh) | 2018-09-30 | 2018-11-30 | 靖西海越农业有限公司 | 用于防治沃柑病虫害的智慧农业监控预警系统 |
| CN110363103B (zh) * | 2019-06-24 | 2021-08-13 | 仲恺农业工程学院 | 虫害识别方法、装置、计算机设备及存储介质 |
| SE545381C2 (en) * | 2019-09-05 | 2023-07-25 | Beescanning Global Ab | Method for calculating the deviation relation of a population registered on image for calculating pest infestation on a bee population |
| EP3798899A1 (de) * | 2019-09-30 | 2021-03-31 | Basf Se | Quantifizierung von pflanzenbefall durch schätzung der anzahl der insekten auf blättern durch neuronale faltungsnetze, die dichtekarten bereitstellen |
| US11425852B2 (en) * | 2020-10-16 | 2022-08-30 | Verdant Robotics, Inc. | Autonomous detection and control of vegetation |
-
2019
- 2019-09-30 EP EP19200657.5A patent/EP3798899A1/de not_active Withdrawn
-
2020
- 2020-02-21 EP EP20158881.1A patent/EP3798901A1/de not_active Ceased
- 2020-09-29 AR ARP200102707A patent/AR120119A1/es active IP Right Grant
- 2020-09-29 EP EP20781367.6A patent/EP4038542B1/de active Active
- 2020-09-29 CN CN202080068396.7A patent/CN114467122A/zh active Pending
- 2020-09-29 ES ES20781367T patent/ES3010674T3/es active Active
- 2020-09-29 BR BR112022002456A patent/BR112022002456A2/pt unknown
- 2020-09-29 CA CA3150808A patent/CA3150808A1/en active Pending
- 2020-09-29 US US17/761,849 patent/US12073327B2/en active Active
- 2020-09-29 WO PCT/EP2020/077197 patent/WO2021063929A1/en not_active Ceased
-
2021
- 2021-02-19 ES ES21705986T patent/ES3032897T3/es active Active
- 2021-02-19 CN CN202180014801.1A patent/CN115104133A/zh active Pending
- 2021-02-19 CA CA3172345A patent/CA3172345A1/en active Pending
- 2021-02-19 WO PCT/EP2021/054233 patent/WO2021165512A2/en not_active Ceased
- 2021-02-19 EP EP21705986.4A patent/EP4107656B1/de active Active
- 2021-02-19 US US17/799,829 patent/US12566959B2/en active Active
- 2021-02-19 BR BR112022016566A patent/BR112022016566A2/pt unknown
Also Published As
| Publication number | Publication date |
|---|---|
| US20230071265A1 (en) | 2023-03-09 |
| US20230351743A1 (en) | 2023-11-02 |
| EP4038542A1 (de) | 2022-08-10 |
| US12566959B2 (en) | 2026-03-03 |
| BR112022002456A2 (pt) | 2022-05-03 |
| US12073327B2 (en) | 2024-08-27 |
| ES3010674T3 (en) | 2025-04-04 |
| WO2021063929A1 (en) | 2021-04-08 |
| BR112022016566A2 (pt) | 2022-12-20 |
| WO2021165512A2 (en) | 2021-08-26 |
| AR120119A1 (es) | 2022-02-02 |
| EP4107656B1 (de) | 2025-04-09 |
| ES3032897T3 (en) | 2025-07-28 |
| CN115104133A (zh) | 2022-09-23 |
| EP3798899A1 (de) | 2021-03-31 |
| WO2021165512A3 (en) | 2021-10-14 |
| CA3172345A1 (en) | 2021-08-26 |
| CN114467122A (zh) | 2022-05-10 |
| EP3798901A1 (de) | 2021-03-31 |
| EP4038542B1 (de) | 2024-11-06 |
| CA3150808A1 (en) | 2021-04-08 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| EP4107656B1 (de) | Quantifizierung des pflanzenbefalls durch schätzung der zahl der insekten auf blättern, durch neuronale faltungsnetzwerke, die trainingsbilder verwenden, die von einem halbüberwachten ansatz aufgenommen werden | |
| US11636701B2 (en) | Method for calculating deviation relations of a population | |
| Gao et al. | Cross-domain transfer learning for weed segmentation and mapping in precision farming using ground and UAV images | |
| Yang et al. | A detection method for dead caged hens based on improved YOLOv7 | |
| EP4256465B1 (de) | System und verfahren zur bestimmung der schädigung von pflanzen nach herbizidanwendung | |
| Ananthajothi et al. | Hybrid heuristic optimization of deep learning models for robust plant disease detection | |
| Mahenthiran et al. | Smart pest management: an augmented reality-based approach for an organic cultivation | |
| Gao et al. | Transferring learned patterns from ground-based field imagery to predict UAV-based imagery for crop and weed semantic segmentation in precision crop farming | |
| Choudhury | Segmentation techniques and challenges in plant phenotyping | |
| Chauhan et al. | Potato disease detection and recognition using deep learning | |
| Omer et al. | An image dataset construction for flower recognition using convolutional neural network | |
| CN117321570A (zh) | 基于从图像中提取的解剖组成部分的掩模而对蚊虫幼虫进行分类的系统和方法 | |
| Sundaram et al. | Detection and Providing Suggestion for Removal of Weeds Using Machine Learning Techniques | |
| Velte | Semantic image segmentation combining visible and near-infrared channels with depth information | |
| Oppenheim et al. | Object recognition for agricultural applications using deep convolutional neural networks | |
| Mumtaz et al. | Robust Machine Learning Approach to Plant Species Classification for Sustainable Agriculture | |
| Conrady | Automated detection and classification of red roman in unconstrained underwater environments using Mask R-CNN | |
| Saini et al. | Artificial Intelligence based Plant Disease Detection System | |
| Jayachandran | Detection and Providing Suggestion for Removal of Weeds Using Machine Learning Techniques | |
| Zolikha | Forest Firefighting Using Drone And Artificial Intelligence | |
| Yuvaraju et al. | A Deep Multi-View Learning Approach for Wildlife Animal Detection and Identification | |
| Minakshi | Automating the Classification of Mosquito Specimens Using Image Processing Techniques | |
| Mitra¹ et al. | Detection and Leaf Damage Estimation | |
| Mitchell | Locust Instar Classification Using Deep Learning | |
| He | Deep learning applications to automate phenotypic measurements on biodiversity datasets |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: UNKNOWN |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE INTERNATIONAL PUBLICATION HAS BEEN MADE |
|
| PUAI | Public reference made under article 153(3) epc to a published international application that has entered the european phase |
Free format text: ORIGINAL CODE: 0009012 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: REQUEST FOR EXAMINATION WAS MADE |
|
| 17P | Request for examination filed |
Effective date: 20220921 |
|
| AK | Designated contracting states |
Kind code of ref document: A2 Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC MK MT NL NO PL PT RO RS SE SI SK SM TR |
|
| DAV | Request for validation of the european patent (deleted) | ||
| DAX | Request for extension of the european patent (deleted) | ||
| REG | Reference to a national code |
Ref country code: DE Ref legal event code: R079 Free format text: PREVIOUS MAIN CLASS: G06K0009000000 Ipc: G06V0010440000 Ref document number: 602021028858 Country of ref document: DE |
|
| GRAP | Despatch of communication of intention to grant a patent |
Free format text: ORIGINAL CODE: EPIDOSNIGR1 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: GRANT OF PATENT IS INTENDED |
|
| RIC1 | Information provided on ipc code assigned before grant |
Ipc: G06N 3/045 20230101ALI20241004BHEP Ipc: G06N 3/082 20230101ALI20241004BHEP Ipc: G06V 20/10 20220101ALI20241004BHEP Ipc: G06V 10/764 20220101ALI20241004BHEP Ipc: G06V 10/82 20220101ALI20241004BHEP Ipc: G06V 10/56 20220101ALI20241004BHEP Ipc: G06V 10/44 20220101AFI20241004BHEP |
|
| INTG | Intention to grant announced |
Effective date: 20241016 |
|
| GRAS | Grant fee paid |
Free format text: ORIGINAL CODE: EPIDOSNIGR3 |
|
| P01 | Opt-out of the competence of the unified patent court (upc) registered |
Free format text: CASE NUMBER: APP_67201/2024 Effective date: 20241219 |
|
| GRAA | (expected) grant |
Free format text: ORIGINAL CODE: 0009210 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE PATENT HAS BEEN GRANTED |
|
| AK | Designated contracting states |
Kind code of ref document: B1 Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC MK MT NL NO PL PT RO RS SE SI SK SM TR |
|
| REG | Reference to a national code |
Ref country code: GB Ref legal event code: FG4D |
|
| REG | Reference to a national code |
Ref country code: CH Ref legal event code: EP |
|
| REG | Reference to a national code |
Ref country code: DE Ref legal event code: R096 Ref document number: 602021028858 Country of ref document: DE |
|
| REG | Reference to a national code |
Ref country code: IE Ref legal event code: FG4D |
|
| REG | Reference to a national code |
Ref country code: ES Ref legal event code: FG2A Ref document number: 3032897 Country of ref document: ES Kind code of ref document: T3 Effective date: 20250728 |
|
| REG | Reference to a national code |
Ref country code: NL Ref legal event code: MP Effective date: 20250409 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: NL Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| REG | Reference to a national code |
Ref country code: AT Ref legal event code: MK05 Ref document number: 1784240 Country of ref document: AT Kind code of ref document: T Effective date: 20250409 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: FI Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 Ref country code: PT Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250811 |
|
| REG | Reference to a national code |
Ref country code: LT Ref legal event code: MG9D |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: GR Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250710 Ref country code: NO Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250709 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: PL Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: BG Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: HR Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: AT Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: RS Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250709 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: IS Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250809 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: LV Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| REG | Reference to a national code |
Ref country code: DE Ref legal event code: R097 Ref document number: 602021028858 Country of ref document: DE |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: SM Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 Ref country code: DK Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: CZ Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: EE Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: SK Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20250409 |
|
| PLBE | No opposition filed within time limit |
Free format text: ORIGINAL CODE: 0009261 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: NO OPPOSITION FILED WITHIN TIME LIMIT |
|
| REG | Reference to a national code |
Ref country code: CH Ref legal event code: L10 Free format text: ST27 STATUS EVENT CODE: U-0-0-L10-L00 (AS PROVIDED BY THE NATIONAL OFFICE) Effective date: 20260218 |
|
| 26N | No opposition filed |
Effective date: 20260112 |
|
| PGFP | Annual fee paid to national office [announced via postgrant information from national office to epo] |
Ref country code: ES Payment date: 20260317 Year of fee payment: 6 |
|
| PGFP | Annual fee paid to national office [announced via postgrant information from national office to epo] |
Ref country code: DE Payment date: 20260220 Year of fee payment: 6 |
|
| PGFP | Annual fee paid to national office [announced via postgrant information from national office to epo] |
Ref country code: IT Payment date: 20260220 Year of fee payment: 6 |
|
| PGFP | Annual fee paid to national office [announced via postgrant information from national office to epo] |
Ref country code: FR Payment date: 20260227 Year of fee payment: 6 |