The disclosure generally relates to image processing by computers, and more in particular relates to techniques for quantifying objects on parts of the plant (or “plant parts” in short). Even more in particular, the disclosure relates to techniques for quantifying plant infestation by estimating the number of insects or other biological objects on plant leaves.
It is well known that agricultural plants—such as crops—grow in environments in that they co-exist with biological objects. These objects tend to be located at the plant, usually by being attached to the plants (at least temporarily); and the objects interact with the plant.
Different objects tend to attach themselves to different parts of the plants. For example, some insects may sit on the leaves, some others may stick to the stem or to a branch, and so one.
From the perspective of the farmer, the object-and-plant interaction has two directions or aspects. In a first direction, there are biological objects with a detrimental effect on the plants. For example, animal pests destroy crops, causing large economic loss to the food supply and to property. To give an illustrative example, a pest animal may eat from the leaves or may eat from the fruits of the crop.
In the other direction, there are also animals to the benefit of the plant. For example, the plant may have a direct benefit when a butterfly visit its flower or blossom, or the plant may have an indirect benefit if a ladybug eats aphids (or other pest).
The farmer may control the presence of these objects: to favor the absence of some objects (such as pest animals) and the presence of beneficiary objects.
The control measures are usually adapted to the quantity of the objects on the plants. In principle it does not matter where the object is located on the plant (leaf, stem, fruit or wherever) and what direction the interaction has.
To illustrate some of these aspects by way of example, insects of many species live on plant leaves. For example, whiteflies live on the leaves of eggplants.
In the broadest sense, the insects interact with the plant (for example by consuming part of the leaves). The insects can cause diseases or other abnormal conditions of the plant. Eventually, the plant does not survive the presence of the insects. But in agriculture, the plants should become food (i.e. crop for humans or animals), and insects being present on leaves are not desired at all. Food security is of vital importance.
Usual terms for such phenomena are “infestation” and “pest”. Farmers apply countermeasures (e.g., treatment by applying insecticides) in order to remove the insects.
However, applying countermeasures may cause further problems or challenges. Countermeasures must be specific to particular insects, for example to remove the whiteflies but to keep the bees and others. Countermeasures should also take the quantity of the insects into account.
Quantifying the infestations, such as by counting insects (on plant leaves) is therefore an important task for pest management.
In theory, farmers could visually inspect the plant and could count the insects (taking the insect development stages into account). As different people have different knowledge (regarding insects) and have different eyes, different people would arrive at different numbers.
Using computer vision techniques appears as an improvement. A well-known (classical or traditional) approach is the extraction of image features with subsequent classification. However, there are many constraints arising. The constraints have many aspects, such as limitations of the computers and cameras, non-ideal conditions in the field and constraints related to the insects themselves.
US 2018/0121764 A1 explains an approach to selectively sterilizing insects. Insects are being reared, placed on a surface, photographed and selectively manipulated by robots. A computer processes the image and identifies location data for particular inspects, and thereby differentiates male insects from female insects. With the location data, the robot can then perform actions with respect to insects, such as removing particular insects.
Xiaoping Wu, Chi Zhan, Yukun Lai, Ming-Ming Cheng and Jufeng Yang: IP102: A Large-Scale Benchmark Dataset for Insect Pest Recognition, IEEE CVPR, pages 8787-8796, 15 Jun. 2019. In this article, Wu et al explain a large-scale database for insect pest recognition.
The constraints are addressed by a computer system, a computer-implemented method and a computer program product for quantifying biological objects on plant parts (in the example: quantifying plant infestation, estimating the number of insects on leaves of a plant).
The computer program product—when loaded into a memory of a computer and being executed by at least one processor of the computer—performs the steps of the computer-implemented method.
In a production phase, the computer applies convolutional neural networks that had been trained previously in a training phase. The production phase is summarized first:
The computer receives a plant-image taken from a particular plant. The plant-image shows at least one of the leaves of the particular plant, the so-called main leaf (or main part).
The computer uses a first convolutional neural network to process the plant-image to derive a leaf-image being a contiguous set of pixels that show a main part of the particular plant completely (i.e. as a whole). The first convolutional neural network has been trained by a plurality of leaf-annotated plant-images, wherein the plant-images had been annotated to identify main parts.
The computer splits the leaf-image into a plurality of tiles. The tiles are segments or portions of the plant-image having pre-defined tile dimensions.
The computer uses a second convolutional neural network to separately process the plurality of tiles to obtain a plurality of density maps having map dimensions that correspond to the tile dimensions. The network having been trained by processing object-annotated plant-images, and the training comprised the calculation of convolutions for each pixel based on a kernel function, leading to density maps. The density maps have different integral values for tiles showing biological objects and tiles not showing biological objects.
The computer combines the plurality of density maps to a combined density map in the dimension of the leaf-image, and integrates the pixel values of the combined density map to an estimated number of biological objects for the part of the plant.
The first convolutional neural network—that is the network to identify the main leaf—can be of the DenseNet type. The second convolutional neural network—that is the network to estimate the number of biological objects on the part—can be of the convolutional neural network (FCRN) type.
The second convolutional neural network (FCRN type) can be a modified network that uses global sum pooling instead of global average pooling.
The convolutional neural network can be also be modified by using dropout.
The second convolutional neural network can also be modified by implementing an input layer as a pixel value filter for individual pixels or for pixel pluralities, the so-called tile segments.
For the pixel pluralities, the network uses a layer to convolute the pixel pluralities, to encode the convoluted segments into segment values so that a further layer applies filtering to the segment values. The second convolutional neural network can also been implemented with a layer to subsequently decode the segment values to the convoluted segments. This approach lets the network layers between the encoder and the decoder operate with numerical values for the segments and not with numerical values for the pixels. Since the number of segments is less than the number of pixels, the approach uses less computation resources (compared to pixel processing).
The second convolutional neural network can be trained by processing object-annotated plant-images for different classes. Classes are identified by object species (e.g., insect species) and by growing stages of the objects. Processing is separated for the classes by branches and output channels.
This approach can be advantageous—in the scenario that the biological objects are pest because it quantifies the plant infestation (by pest) with a granularity that allows fine-tuning countermeasures to particular pest classes.
In the production phase, receiving the plant-image can be performed by receiving the plant-image from the camera of a mobile device. This can be advantageous because mobile devices are readily available to farmers and because the mobile device allows the immediate communication of the plant-image to the computer that quantifies the plant infestation.
Receiving the plant-image can comprise evaluating the class (e.g., the pest class) and the pixel resolution of the camera of the mobile device according to pre-defined rules, wherein for some classes and resolutions, the mobile device is caused to take the image with a magnifying lens. This measure can improve the accuracy of the estimation. The farmer can be instructed via the user interface of the mobile device to apply the lens.
The second convolutional neural network can have been trained by using a loss-function being the mean absolute error or being the mean square error.
The description starts by explaining some writing conventions.
The term “image” stands for the data-structure of a digital photograph (i.e., a data-structure using a file format such as JPEG, TIFF, BMP, RAW or the like). The phrase “take an image” stands for the action of directing a camera to an object (such as a plant, or a part of a plant) and letting the camera store the image.
The description uses the term “show” when it explains the content of images (i.e., the semantics), for example in phrases such as “the image shows a plant”. There is however no need that a human user looks at the image. Such computer-user interactions are expressed with the term “display”, such as in “the computer displays the plant-image to an expert”, where an expert user looks at the screen to see the plant on the image.
The term “annotation” stands for meta-data that a computer receives when an expert user looks at the display of an image and interacts with the computer. The term “annotated image” indicates the availability of such meta-data for an image (or for a sub-region of that image), but there is no need to store the meta-data and the image in the same data-structure. Occasionally, the drawings illustrate annotations as part of an image, such as by polygon and/or dots, but again: the annotations are meta-data and there is no need to embed them into the data-structure of the image.
The images will only show plants with their above-ground (air) components, but not the root. Therefore, the term “plant part” (or “part” in short) refers to any of the following: stem, branch, leaf, flower (or blossom), fruit, bud, seed, fruit, node, and internode. The same principle applies to the plural form “parts”: stems, branches, leaves, flowers and so on. Of course, not every plant will have parts in each category, so a young plant may not yet have fruits. As used herein, the description writes “leaf” as pars per to for “part”.
This convention also applies to phrase such as “leaf-annotated” standing for “part-annotated”.
In general, the term “insect” stands for animalia in the phylum “Arthropoda” or 1ARTHP (EPPO-code by the European and Mediterranean Plant Protection Organization). In implementations, the insects are of the subphylum “Hexapoda” (1HEXAQ). The description uses the term “insect” for simplicity and for convenience. It is noted that “insect” is a noun (in usual language) that most readers can easily apply for counting. The skilled person reading “one insect” or “two insects” immediately understands.
The term “insect” is also used to represent biological objects that are located on parts of the plant be counted.
A biological object (to be counted) has a physical size that is relatively smaller than the part on that it is located. It is also noted that the objects are located on one part. Since the plant images are processed to images showing one part (by segmentation), the size relation also transfers to the image.
To illustrates that: an insect sitting with some legs on a first leaf, and sitting with the other legs on a second leaf is not counted because the image would be segmented to one of the leaves. Or, a relatively large insect that shows up on the image covering a leaf and covering a branch could not be counted.
In terms of biological taxonomy, the biological objects can be insects or can be arachnids (i.e., being arthropoda), or the biological objects can be mollusca (not arthropod).
The internal structure of the biological objects does not matter, as long as it fits the size criterion. On other words, it does not matter is the object has an exoskeleton, a segmented body, and paired jointed appendages (as arthropods have) or not.
Further, for the computer, the different number of legs (e.g., insects 6 legs, arachnids 8 legs, or even no legs as with snails) does not matter for the computer.
The objects can also be spots on the surface of the plant parts (spots that are the result of biological processes, such as fungi interacting with the plant or the like, animal excrements, etc.). The person of skill in the art can identify suitable measures (e.g., countermeasures).
The use of the term “insects” is applicable to phrases such as “insect-annotated” or the like to the meaning “object-annotated”.
Further, the interaction of the biological object with the plant (or with the stem branch etc. part) does not matter. The biological object can be pest or beneficial. The term “stage” (also “development stage”, “growing stage”) identifies differences in the life-cycle (or metamorphosis) of insects (i.e., of the biological objects in general), wherein an insect in a first stage has a different visual appearance than an insect in a second, subsequent stage. Biologists can differentiate the stages (or “stadia”) by terms such as egg, larva, pupa, and imago. Other conventions can also be used, such as “adults” and “nymphs”, or even standardized numerical identifiers such as “n1n2”, “n3n4” and so on. Development stages of the plants are not differentiated.
The term “count” is short for “estimating a number”, such as for estimating the number of insects on a leaf (i.e., number of biological objects on plant parts).
The description uses the term “train” as a label for a first process—the “training process”—that enables CNNs to count insects, and for the particular task to train a particular CNN by using annotated images.
For convenience, the description refers to hardware components (such as computers, cameras, mobile devices, communication networks) in singular terms. However, implementations can use multiple components. For example, “the camera taking a plurality of images” comprises scenarios in that multiple cameras participate so that some images are taken from a first camera, some image are taken from a second camera and so on.
In the figures, the suffixes “−1, −2 . . . ” and so on distinguish like items; and suffixes “(1), (2) . . . ” distinguish different items.
The term “class” is used in the general meaning to differentiate sets or the like, in case that the term “class” refers to a taxonomic rank in biology, this will be explained if needed.
Referring to
Throughout this description, references noted as **1/**2 stand for elements that are similar but that have different use in both phases.
From left to right,
Some of the auxiliary activities are pre-processing activities that prepare method executions. In
Computers 201/202 use CNNs and other modules to be explained below (such as user interfaces, databases, splitter and combiner modules etc.). While
Methods 601B and 602B are performed with CNNs 261/262, and methods 701B and 702B are performed with CNNs 271/272. CNNs 261/262 and CNNs 271/272 differ from each other by parameters (explained below).
The CNNs use density map estimation techniques, where—simplified—the integral of the pixel values leads to the estimated insect numbers. In other words, counting is performed by calculating an integral. The estimated numbers NEST can be non-integer numbers. For the above-mentioned purpose (to identify appropriate countermeasures against the infestation), the accuracy of NEST is sufficient.
Using density maps to count objects is explained by “Lempitsky, V., Zisserman, A., 2010. Learning To Count Objects in Images. Neural Inf. Process. Syst. 1-9.”
Training phase **1 is illustrated in the first row of
As illustrated by pre-processing 601A, camera 311 takes a plurality of plant-images 411 (in an image acquisition campaign). Computer 301 interacts with expert user 191 to obtain leaf-annotations and to obtain insect-annotations. User 191 can have different roles (details in
Computer 301 forwards annotated images 461, 471 to computer 201.
In performing computer-implemented method 601B, computer 201 (
In performing method 701B, computer 201 receives the plurality of leaf-annotated plant-images in combination with insect-annotations (collectively “insect-annotated leaf-images”). Computer 201 then trains CNN 271 to count insects on particular leaves. Thereby, computer 201 turns un-trained CNN 271 into trained CNN 272. In other words, CNN 272 is output of computer 201 as well.
It is noted that the description assumes the annotations to be made for the same plurality of plant-images 411. This is convenient, but not required. The pluralities can be different. For example, the plurality of plant-images 411 to be leaf-annotated can show non-infested plants. Using leaf-annotated plant-images 471 (from such healthy plants) to further provide insect-annotations would fail because there would be no insects to annotate. Providing insect-annotations could be performed for images that are not segmented to leaves.
Production phase **2 is illustrated in the second row of
As illustrated by pre-processing 602A, camera 312 of device 302 takes plant-image 412 and forwards it to computer 202.
In performing method 602B, computer 202 (
In scientific literature, using trained CNNs to obtain results is occasionally called “testing”.
The description now explains further aspects and implementation details, again in view of
Training Phase with Details
Returning to
Training phase **1 has two sub-phases.
Although computer 201 is illustrated by a single box, it can be implemented by separate physical computers. The same principle applies for plant 111 and for camera 311. The plant and the camera do not have to be the same for all images. It is rather expected to have plant-images 411 from cameras 311 with different properties. Also, the plurality of images 411 represents a plurality of plants 111. There is no need for a one-to-one relation, so one particular plant may be represented by multiple images.
Training the CNNs can be seen as the transition from the training phase to the production phase. As in
There is no need to copy the CNNs (quasi from figure to figure). The person of skill in the art can take over parameters from one network to another, such as from CNN 261 (of
As in
Training phase **1 is usually performed once, in supervised learning with expert user 191. The setting for the training phase with camera 311 taking plant-images 411 (as reference images), with expert user 191 annotating plant-images 411 (or derivatives thereof) and with computer-implemented processing will be explained. The description assumes that training phase **1 has been completed before production phase **2. It is however possible to perform training phase **1 continuously and in parallel to production phase **2.
Production Phase with Details
Returning to
Simplified, computer 202 processes plant-image 412 received from mobile device 302 through communication network 342. In difference to training phase **1 of
Leaves 122 are so-called “infested leaves” because insects are located on them. Counting can be differentiated for insects of particular class 132(1) (illustrated by plain ovals) and—optionally—of particular class 132(2) (bold ovals). Optionally, counting can consider further classes (cf.
Non-insect objects 142 are not necessarily to be counted. Such objects 142 can be located within the leaf and can be structural elements of leaves 122, such as damages on the leaf, shining effects due to light reflection or the like. It is noted that many insects camouflage themselves. Therefore, for the computer it might be difficult to differentiate insects 132 and non-insect objects 142.
Insect classes (1) and (2) are defined
A more fine-tuned granularity with more classes is given in
Plant 112 has a plurality of leaves 122. For simplicity, only two leaves 122-1 and 122-2 are illustrated. Leaves 122 are occupied by insects 132 (there is no difference to the training phase **1). For convenience,
Mobile device 302 can be seen as a combination of an image device (i.e. camera 312), processor and memory. Mobile device 302 is readily available to the farmers, for example as a so-called “smartphone” or as a “tablet”. Of course, mobile device 302 can be regarded as a “computer”. It is noted that mobile device 302 participate in auxiliary activities (cf.
Field user 192 tries to catch at least one complete leaf (here leaf 122-1) into (at least one) plant-image 412. In other words, field user 192 just makes a photo of the plant. Thereby, field user 192 may look at user interface 392 (i.e. at the visual user-interface of device 302) that displays the plant that is located in front of camera 312.
Mobile device 302 then forwards plant-image 412 via communication network 342 to computer 202. As the illustration of communication network 342 suggests, computer 202 can be implemented remotely from mobile device 302.
Computer 202 returns a result that is displayed to user interface 392 (of mobile device 302). In the much simplified example of this figure, there are N(1)=3 insects of class (1) (i.e., insects 132(1)) and N(2)=2 insects of class (2) (i.e., insects 132(2)) counted. The numbers N(1), N(2) (or N in general) are numbers-per-leaf, not numbers per plant (here in the example for main leaf 122-1). The numbers correspond to NEST (with NEST being rounded to the nearest integer N).
Optionally, by proving infestation data that is separated by classes (such as (1) and (2)), field user 192 can identify countermeasures to combat infestation with better expectation of success.
The term “main leaf” does not imply any hierarchy with the leaves on the plant, but simply stands for that particular leaf for that the insects are counted. Adjacent leaf 122-2 is an example of a leaf that is located close to main leaf 122-1, but for that insects are not to be counted. Although illustrated in singular, plant 112 has one main leaf but multiple adjacent leaves. It is assumed that plant-image 412 represents the main leaf completely, and represents the adjacent leaves only partially. This is convenient for explanation, but not required.
It is usual that main leaf 122-1 is on top of adjacent leaf 122-2 (or leaves). They overlap each other and it is difficult to identify the edges between one from the other.
The numbers N are derived from estimated numbers NEST, the description describes an approach to accurately determine N.
Ideally, only insects 132 located on main leaf 122-1 are counted. Insects 132 that are not counted but that are located on main leaf 122-1 would be considered to be “false negatives”, and insects that are counted but that are located on an adjacent leaf would be considered to be “false positives”. As it will be explained with more detail below, counting comprises two major sub-processes:
In other words, identifying the main leaf prior to counting keeps the number of “false negatives” and “false positives” negligible.
The communication between mobile device 302 and computer 202 via communication network 342 can be implemented by techniques that are available and that are commercially offered by communication providers.
Computer 202 has CNN 262/272 that performs computer-implemented method 602B and 602B (details in connections with
In an embodiment, computers 201/202 use operating system (OS) Linux, and the module that executes methods 601B/602B, 701B/702B was implemented by software in the Python programming language. It is convenient to implement the modules by a virtualization with containers. Appropriate software is commercially available, for example, from Docker Inc. (San Francisco, California, US). In a software-as-a-service (SaaS) implementation, mobile device 302 acts as the client, and computer 202 acts as the server.
Besides CNNs 262/272, computer 202 has other modules, for example, a well-known REST API (Representational State Transfer, Application Programming Interface) to implement the communication between mobile device 302 and computer 202 can use. Computer 202 appears to mobile device 302 as a web-service. The person of skill in the art can apply other settings.
The time it takes computer 202 with CNNs 262/272 (performing the method) to obtain NEST depends on the resolution of plant-image 412. Performing methods 602B and 702B may take a couple of seconds. The processing time rises with the resolution of the image. (The processing time has been measured in test runs. For plant-image 412 with 4000×6000 pixels, the processing time was approximately 9 seconds.)
It is convenient, to transmit plant-image 412 in its original pixel resolution, otherwise the accuracy to count insects will deteriorate. In other words, there are many techniques to transmit images in reduced resolutions, but for this application, such techniques should be ignored here. However, transmitting a compressed image (in a loss-less format) can be possible. In modern communication networks, the bandwidth consumption (for transmitting the image in original resolution) is no problem any longer.
It is noted that for field user 192, the conditions for catching images are not always ideal. For example, there are variations in
In the following, the description shortly investigates the objects to be counted: insects 131/132 (cf.
The description uses two examples of plant/insect combinations.
It is noted that the person of skill in the art can differentiate such (and other combinations) without further explanations herein. Taking images, annotating images, training the CNNs, counting insects (cf. pre-processing and method execution in
Exceptions from the general rule are available. Having different plant species in the training phase **1 and the production phase **2 can be possible if the plants are similar in appearance. In that case pre-processing 601A and executing method 601B (i.e. to train CNN 261/262 to segment leaves) would be performed with a first plant species (e.g., eggplant) and pre-processing 602A and executing method 602B would be performed with a second plant species. The second plant species can belong to other crops such as for example cotton, soy bean, cabbage, maize (i.e., corn).
Since the infestation is made by the insects, the description focuses on the insects. It is a constraint that insects change appearance in the so-called metamorphosis with a sequence of development stages.
As the accuracy in obtaining data regarding infestation is related to the efforts to obtain the data, the description now introduces granularity aspects.
As illustrated by arrows (from left to right), the development stages occur in a predefined sequence with state transitions: from stage (A) to (B), from (B) to (C), from (C) to (D). The arrows are dashed, just to illustrate that other transitions (such as from (B) to (D)) are possible. Biologists can associate the stages with semantics relating to the age of the insects, such as “egg”, “nymph”, “adult”, “empty pupae” (an insect has left the pupa and only the skin of the pupa is left), with semantics relating to life and death. As particular way to express stages is the “n1n2”/“n3n4” nomenclature, well known in the art.
Details for the appearance in each stage are well-known. Just to mention one point, insects can develop wings. For example, the presence or absence of wings can indicate particular development stage for thrips.
Below the stages,
In the example there are two species: (i) “whitefly” and (ii) “thrips”. Insects of both species develop through the (A) to (D) stages (of course separately: (i) do not turn into (ii) or vice versa). The black dots at the column/row crossings indicate that insects of particular stage/species combinations should be counted. This is a compromise between accuracy (e.g. infestation critical for black dotted situations, but countermeasures available) and efforts (annotations, calculations, training etc.).
Rectangles group the particular stage/species combinations into classes (1) to (4) and thereby differentiate use cases 1 to 3.
For each particular use case, the following assumptions applies:
The description explains use cases by example:
In use case 1, CNN 271/272 is trained to provide NEST as the number of species (i) insects in stages (B) and (C), without differentiating (B) and (C), that is
In use case 2, CNN 271/272 is trained to provide NEST in 2 separate numbers (cf. the introduction in
In use case 3, CNN 271/272 is trained to provide NEST in 4 separate numbers:
The rectangles are illustrated with class numbers (1) to (4), wherein the classes are just alternative notations. The description will explain adaptations to the CNNs for multi-class use cases (use cases 2 and 3) in connection with
The description now refers to some challenges, but in combination with solution approaches.
The impact of the insects to the plant (as well as the appropriate countermeasures) can be different for each development stage. For example, it may be important to determine the number of nymphs (per leaf), the number of empty pupae and so on. Differentiating between young and old nymphs can indicate the time interval that has passed since the arrival of the insects, with the opportunity to fine-tune the countermeasure. For example, adults may lay eggs (and that should be prevented).
It is noted that the spatial arrangement (i.e. pixel coordinates (X, Y)) of the leaves, the insects and the non-insect objects is different from image to image. The reason is simple: the images show different physical plants (even taken at different time points).
The description uses terms such as “insect 431/432” and “non-insect object 441/442” for convenience of explanation. It is however noted that
For use in training phase **1, plant-image 411 (cf.
Although illustrated here as a single image, in training phase **1, images are taken in pluralities. It is noted that the variety of different cameras can be taken into account when taking images for training.
In the production phase, plant-image 412 is usually taken by camera 312 of mobile device 302 (cf.
Image 411/412 is usually a three-channel color image, with the color usually coded in the RGB color space (i.e., red, green and blue).
It is noted that image 412 does not have to be displayed to field user 192. Also, the field scenario will be explained for a single image 412, but in practice it might be advisable for field user 192 to take a couple of similar images 412.
Image 412 represents reality (i.e. plant 112, leaves 122, insects 132, non-insect objects 142), but with at least the following further constraints.
As mentioned already, plant 111/112 has multiple leaves at separate physical locations. Therefore in image 411/412, one leaf can overlay other leaves. Or in other words, while in reality (cf.
However, with the goal to count N as “insect in a particular class per leaf”, the overlay must be considered. As multiple leaves have similar color (usually, green color), their representations in plant-image 411/412 appear in the same color (i.e., small or zero color difference in the image).
Further, each insect of a particular class has a particular color. This color could be called text-book color, or standard color. For example, as the name suggests, an adult whitefly is white (at least in most parts).
However, the image would not properly represent the text-book color. There are at least the following reasons for that:
The illumination can be different (e.g., cloudy sky, sunny sky, shadow and so on)
In the coding of the image (the numerical values that represent color, e.g., in the mentioned RGB space), the numerical values would be different. Therefore, the absolute value (of the color) in the image is therefore NOT particular characterizing.
Due to the mentioned camouflage, it can be complicated to differentiate insect 131/132 from non-insect object 141/142. This is complicated in nature and even more complicated in images.
Further, insects can be relatively tiny in comparison to the leaves. For example, an insect can be smaller than one millimeter in length. In contrast to the emphasis in
Further, it is natural behavior of the insects to sit on the leaf close to each other. In other words, insects tend to be present on the leaf in pairs (i.e., two insects), or even in triples (i.e., three insects). So in other words, a 30×20 pixel portion of image 411/412 might represent two or more insects.
The pixel numbers 20×30 are exemplary numbers, but it can be assumed that insects 431/432 are dimensioned with two-digit pixel numbers (i.e. up to 99 pixels in each of the two coordinates). The same limitations can be true for non-insect objects 441/442.
As it will be explained, CNNs 271/272 (to count insects) use density map estimation (instead of the above-mentioned traditional object detection). In density maps, insects would be represented as areas, and the integral of the pixel values of the area would be approximately 1 (assuming that the pixel values are real numbers, normalized between 0 and 1, and also assuming to have one insect per map). It is noted that for situations in that two insects are located close together and overlapping on the image, there would be a single area, but the sum of the pixel values would be approximately 2.
It is noted that insects of two or more stages can be available on a single leaf at the same time. It is a constraint that the differences between two stages can be subtle. For example, on a leaf in reality, insects in stages (C) and (D) may look similar.
As a consequence, a computer using a conventional computer-vision technique (such as the mentioned technique with feature extraction) may not recognize the differences. However, an expert user can see differences (on images), and training images can be properly annotated (cf. use case 1).
There are also constraints related to mobile device 302 (cf.
Further, the farmer (i.e. the user of the mobile device) requires a result shortly after taking the image. To be more accurate: the time interval from taking the image to determining the insect-number-per-leaf must be negligible so that
The identification and the application of the countermeasures can only start when the insect-number-per-lead has been established. A countermeasure—although properly identified—may be applied too late to be effective. For example, a countermeasure that is specialized to destroy eggs would not have any effect if the insects have already hatched from the eggs (cf. stage specific countermeasures).
The following is taken into account by the solution. The species of the plant is usually known (for example, the farmer knows eggplant) so that the computer has the information (as an attribute of the image). Therefore, the plant species is therefore not further discussed here.
Those of skill in the art can implement the interaction between computer 301 and expert user 191 by appropriate user interfaces, for example with a display showing images and with interface elements to identify parts of the image (e.g., touch-screen, mouse, keyboard etc.). Software tools for such and other annotations are known in the art. A convenient tool “LabelMe” is described by Russel, B. C., Torralba, A., Murphy, K. P., Freeman, W. T., 2008. 2008 LabelMe. Int. J. Comput. Vis. 77, 157-173. doi:10.5591/978-1-57735-516-8/IJCAI11-407.
Expert user 191 conveys ground truth information to the images, not only regarding the presence or absence of a main leaf (by the leaf-annotations), or the presence or absences of particular insects (by the insect-annotations), but also information regarding the position of the main leaf and of the insect in terms of (X, Y) coordinates. Depending on the selected granularity of the use cases (cf.
Both annotation processes (leaf annotations, insect annotations) can be performed independently, and even the expertise of user 191 can be different. For both stages, user 191 rather assumes particular roles:
Although
The description now explains details for each type of annotation separately:
As illustrated on the left side of
The leaf-annotation allows computer 201 (cf.
For the leaf-annotation, it does not matter if the leaf shows insects (or non-insect objects).
As illustrated on the right side of
The insect annotation also identifies the position of the insects (and/or non-insect objects) by coordinates.
It is noted that the annotations can take the use cases (cf.
Expert user 191 can actually set the dots next (or above) to the insects. As used herein, a single dot points to a particular single pixel (the “dot pixel” or “annotation pixel”). The coordinate of that single pixel at position coordinate (X′, Y′) of an insect (or non-insect object) is communicated to computer 301.
Computer 301 stores the position coordinates as part of the annotation. Coordinates (X′, Y′) can be regarded as annotation coordinates, and the computer would also store the semantic, such as (i)(C) in annotation 1, as (i)(B) in annotation β and so on.
As it will be explained further, the insect-annotation (for a particular image 411) is used by computer 201 in training CNN 271, for example, by letting the computer convolute images (i.e., tiles of images) with kernel functions that are centered at the position coordinate (X′, Y′). Also, the insect-annotation comprises ground truth data regarding the number of insects (optionally in the granularity of the use cases of
The annotations can be embedded in an annotated image by dots (in color coding, e.g. red for stage (C), stage (D), or as X, Y coordinates separately.
Using dot annotations is convenient, because the (X, Y) coordinates of the annotations indicate where the insects are shown on the image.
In training phase **1, computer 201 obtains leaf-image 421 through interaction with expert user 191, as explained below (cf.
In the production phase, computer 202 obtains leaf-image 422 through segmenting plant-image 412 by using (trained) CNN 262 (in method 602B). In the production phase, annotations are not available. It is noted that leaf-image 422 (production phase) is not the same as leaf-image 421 (training phase).
Reference 429 illustrates portions of leaf image 421/422 that do not show the main leaf. The pixels in portions 429 can be ignored in subsequence processing steps. For example, a processing step by that an image is split into tiles does not have to be performed for portions 429 (because insects are not to be counted according to the insects per leaf definition). In implementations, these portions 429 can be represented by pixels having a particular color or otherwise. In illustrations (or optionally in displaying portions 429 to users), the portions can be for example displayed in black or white or other single-color (e.g., white as in
The number of tiles 401-k/402-k in image 401/402 is given be reference K. In the example, image 400 can have an image dimension of 4000×6000 pixels (annotations do not change the dimension). The tiles have tile dimensions that are smaller than the image dimensions. For example, the tile dimension is 256×256 pixels. The tile dimensions correspond to the dimension of the input layer of the CNNs (cf.
The figure illustrates particular tiles 401-k/402-k in a close-up view on the right side, with examples:
In the example alpha, tile 401-k was split out from annotated image 471. Therefore, annotations are applicable for tile 401-k as well. For example, if an annotation indicates the presence of insect 431 for a particular (X′, Y′) coordinate, tile 401-k comprises the pixel with that particular coordinate and tile 401-k takes over this annotation (cf. the dot symbol, with position coordinates (X′, Y′) cf.
During training, CNN 271 would learn parameters to obtain density map 501-k with the integral summing up to 1 (corresponding to 1 insect, assuming normalization of the pixel values in the density maps). For example, CNN 271 would take the position coordinate (X′, Y′) to be the center for applying a kernel function to all pixels of tile 401-k.
In the example beta, tile 402-k was split from image 412 (production phase), it shows insect 432. Of course, the insect is not necessarily at the same position as in “annotated” tile 401-k above in alpha). Using the learned parameters, CNN 272 would arrive at density map 502-k with integral 1.
The example gamma is a variation of the example alpha. Tile 401-k was split up, and annotations are de facto available as well. Although expert user 192 did not provide annotations (dots or the like), the meta-data indicates the absence of an insect. During training, CNN 271 would learn parameters to obtain density map 501-k with the integral summing up to 0.
The example delta is a variation of case beta. A non-insect tile 402-k is processed in the production phase (by CNN 272) and it would arrive at a density map with integral 0.
It is noted that—in the production phase—the density maps are provided for all tiles 402-k (k=1 to K). The person of skill in the art can implement this, for example, by operating CNN 272 in K repetitions (i.e. one run per tile), and combiner module 282 can reconstruct the density maps of the tiles in the same order as splitter module 242 has split them (cf.
Taking the Use Cases and the Classes into Account
While in
As explained above, expert user 192 can annotate images for different insect species (e.g., (i) and (ii)), development stages (for example (A) to (B), at least for the combinations highlighted in
Taking use case 3 as an example, there can be annotations for (i) (A), (i) (B), (i) (C), and (i) (D) (i.e. whitefly in four stages, classes (1) to (4)). In other words, the annotations are class specific.
This leads to different tiles 401-k for these combinations (or classes). Density maps 501-k (training phase) and 502-k (production phase) are different as well, simply because images 411/412 are different. The integrals can be separately calculated for different insect classes, resulting in the NEST (for the complete image, after combination) specific for the classes (cf.
Splitter module 241/242 receives images 411/412 (images 411 with annotations as images 461, 471) and provides tiles 401-k/402-k. As it will be explained, the CNNs provide density maps, combiner module 281/282 receives maps 501-k/502-k and provides combined density map 555.
Since there is an overlap (cf.
Combiner module 281/282 can also calculate the overall integral of the pixel values (of the combined density map), thus resulting in NEST.
As already mentioned in connection with
The CNNs do not receive the images in the original image dimension (e.g., 4000×6000 pixels) but in tile/map dimensions (e.g., 224×224 pixels).
Density Map calculation
During processing, the CNNs obtain intermediate data. For convenience,
There is however no need to display the tiles and the maps to a user.
Map 502-k is a density map derived from tile 402-k. Map 502-k has the same dimension as the tile 402-k. In other words, the map dimensions and the tile dimensions are corresponding to each other. The density map can be understood as a collection of single-color pixels in X-Y-coordinates, each having a numerical value V(X, Y). The integral of the values V of all X-Y-coordinates corresponds to the number of objects (i.e. insects). In the example, the integral is 2 (in an ideal case), corresponding to the number of insects (e.g., two insects shown in tile 402-k).
In the production phase **2, map 502-k is obtained by prediction (with the parameters obtained during training). During the training phase **2, one of the processing steps is the application of a kernel function (e.g., a Gaussian kernel) with the kernel center corresponding to an annotation coordinate (X′, Y′), if an annotation (for an insect) is available in the particular tile 402-k. In other words, during training the tiles with annotations are processed to normalized Gaussians. In the absence of annotations, kernel functions are not applied.
Since tile 402-k is only a portion of the (complete) image (at the input of splitter 242), combiner module 282 (cf.
Converting tile 401-k to map 501-k is based on layer-specific parameters obtained by training (i.e., training CNN 271 to become CNN 272). Since the insect-annotations (cf.
In the example, tile 401-k has the annotation “2 insects”. It is noted that both insects can belong to different classes (cf.
Networks are publicly available in a variety of implementations, and the networks are configured by configuration parameters.
The description shortly refers to input/output parameters in general as well as to configuration parameter (in connection with
Exemplary networks comprise the following network types (or “architectures”):
The CNNs have the following properties:
The following particulars are introduced (or used) by setting parameters accordingly (skilled person):
The FCRN network (by Xie et al) was modified by the following:
Also (for all 3 network types), the following parameter settings are useful:
Convenient parameters are also the following:
Auxiliary parameters can be used to deal with technical limitations of the computers. For example, computer 201/202 that implements the CNNs may use floating point numbers, with a maximum highest number of 65536. However, numerical values that the CNN uses to decide for activation (non-activation) could be in the range between 0.0000 and 0.0067 (e.g., in Gaussian kernel with σ=9).
It may be problematic that CNN 271/272 is not capable of learning what information has to be learned. This is because the contrast (in a density map) between insect (pixel activation of 0.0067) and “no insect” (pixel activation of 0.00) is relatively small. Applying a scale factor increases the contrast in the density maps, and eases the density map estimations (i.e., with integrals over images indicating the number of objects). The scale factor can be introduced as auxiliary parameter. For example, all pixel values may be multiplied by the factor 50.000 at the input, and all output values (i.e., insect counts) would be divided by that factor at the output. The factor just shifts the numerical value into a range in that the computer operates more accurately. The mentioned factor is given by way of example, the person of skill in the art can use a different one.
In implementations, CNN 261/262 (to detect leaves) is a CNN of the DenseNet type. For this purpose, the following parameters are convenient:
In implementations, CNN 271/272 (i.e. the CNN to detect insects) is a CNN of the FCRN type. For this purpose, the following parameters are convenient:
While
On the left side,
Each tile should have pixels from p=1 to p=P. For tiles with 256×256 pixels, there are 65.536 pixels. The figure illustrates tiles with 12×12=144 pixels just for simplification.
Each pixel “pix” has a RGB triplet (r, g, b) that indicate the share of the primary colors. There are many notations available, for example each share could also be noted by an integer number (e.g. from 0 to 255 for each color in case of 8 bit coding per color).
The filter condition can be implemented, for example, such that pixels from the input are forwarded to the output if the pixel values comply with color parameters Red R(c), Green G(c) and Blue B(c). The conditions can be AND-related.
The color parameters are obtained by training (the insect classes annotated, as explained above).
Much simplified, in a hypothetical example, there should be insects of a first class (first) and of a second class (second). The insects in (first) should be “red”, so that the parameters are R(first)>0.5, G(first)>0.0, and B(first)<0.5. An input pixel that complies with the condition is taken over as an output pixel.
The insects in the (second) class should be “blue”, so that the parameters are R(first)<0.5, G(first)>0.0, and B(first)>0.5. Such an insect is illustrated at the lower part of the input tile, again here much simplified with 3 pixels.
An input pixel that complies with the condition (for class (second)) is taken over as an output pixel, and
It is noted that applying the filter is conveniently implemented as a convolutional layer in CNN 271/272 (the filter filtering tiles, cf.
The filter can be part of the processing channel of CNN 271/272 before the layer(s) that creates the density maps. Therefore, the density maps are class specific.
The color parameters Red R(c), Green G(c) and Blue B(c) are just examples for parameters that are related to pixels, but the person of skill in the art can use further parameters such as transparency (if coded in images) etc.
The figure is simplified, but in implementations, the 256×256 tiles are separated into a different number of pixels, potentially having more pixels per segment.
Segment #10 is being convoluted (with a particular convolution variable, e.g., 3 pixels) to modified segment #10′. Thereby, the pixels values (of the 9 pixels) change.
For example, segment #10 can have the pixel values (1, 0, 1, 0, 1, 0, 0, 0, 0) and segment #10′ can have pixel values (0.8, 0.1, 0.0, 0.1, 0.2, 0.7, 0.1, 0.7, 0.0). As
Filter criteria can now be applied to the modified segment #10′. In this respect,
In other words, CNN 271/272 can then perform subsequent processing steps by using the segment codes (the segment-specific values). This reduces the number of pixels to be processed (simplified, by a factor that corresponds to the number of pixels per segment, with 9 in the illustrative example). In one of the last layers, CNN 271/272 can then apply decoding.
The separation into class-specific layers (by a filter, such as explained in the example of
On the right side,
Density maps 502 that indicate the presence of an insect (in the particular class) are illustrated with a dot. As in the example, density map 502-29 (in the example illustrated as the map with k=29) indicates an insect of class (1) and an insect of class (2).
In the simplified overview, there the overall integral (k=1 to K) for the combined density maps 555 leads to different estimations: NEST (1)=2, NEST (2)=3, NEST (1)=2, NEST (3)=2, and NEST (4)=7 (map 502-9 reflects 2 insects). The overall number of insect is NEST (1)(2)(3)(7)=14. The illustration of
Providing infestation data separated for species and growing stage can be advantageous for the farmer to identify the appropriate countermeasures.
As training is separated for the classes, training is performed separately as well.
It is noted that the illustration as separate branches (i.e., in parallel) is convenient for explanation, but not required. Parallel processing is possible, but the channels can be implemented by serial processing as well. In other words, CNN 271 would be trained for insects in class (1), then for insects in class (2) and so on. In the production phase, CNN 272 would provide density maps for inspects of class (1), then of class (2) and so on.
The description now explains further details regarding the training phase **1 by that CNN 261 is enabled to segment leaves (by becoming CNN 271, method 601B) and enabled to count insects (by becoming CNN 272, method 701B).
As explained above, in production phase **2, CNN 272 provides NEST (the estimated number of insects per leaf for particular plant-image 412) as the output. In an ideal situation, the combination of CNN 262 and CNN 272 would calculate NEST to be exactly the so-called ground truth number NGT: here the number of insects sitting on the particular main leaf of the plant (from that farmer 192 has taken image 412). The difference between NEST and NGT would indicate how accurate camera 312 and CNNs 262/272 are performing.
However, farmer 192 would not manually count the insects (NGT). The description now explains how the accuracy of CNN 262/272 is validated. As insect-annotations identify insects (as leaf-annotations identify leaves), the ground truth numbers NGT are known for annotated images 471 already. That data is used as explained in the following:
The sub-sets have cardinalities S1, S2 and S3, respectively. The number of insect-annotated images S is the sum of the subsets: S=S1+S2+S3.
Differentiating images into sub-sets in known in the art. Therefore,
The cardinalities are specific to the cases (cf.
CNN 261/271 have been trained with the S1 images of the training sub-set to become trained-CNN 262, 272 Trained-CNN 262, 272 have been used to estimate NEST for the S2 images of the validation sub-set. The S2 values NGT are known from the insect-annotations. If for a particular image, NEST is higher than NGT CNNs 262/272 have counted more insects that present in reality.
Below the testing sub-set,
Most of the dots are located approximately along regression line 504. The (graphical) distance of a dot from line 504 indicates the quality of the estimation. Dot 505 stands for an outlier, with much more insects estimated than present.
A metric can be defined as Mean Absolute Error (MAE), or MAE=NEST s−NGT s (The formula given here is simplified, MAE is actually calculated as the sum of the MAEs for s=1 to S3 divided by S3).
Since this is a mean value, NEST and NGT are obtained as the average of the S3 images.
A further metric can be defined as Mean Square Error (MSE), or MSE=ROOT [(NEST−NGT)2]. Again, NEST AND NGT for all S3 has to be taken into account (i.e. [ ] being the sum of ( )2 for all S3).
For case 1 (whitefly single class (1)), the set S3 provided values MAE=3.4 and MSE=7.8. In comparison to a traditional approach (candidate selection with subsequent classification, MAE=8.6 and MSE=11.4), the error values are smaller. In other words, the error by the new approach is less than half of the error of the traditional error.
Differentiating the main leaf from its adjacent leaves (or neighbor leaves) can be implemented by known methods as well (among them feature extraction). For leaves that are green over a non-green ground, color can be used as a differentiator. However, such an approach would eventually fail for “green” over “green” situations, for example, when one leaf overlaps another leaf.
In alternative implementations, counting insects can be implemented by other known approaches, such as by the above-mentioned candidate selection with subsequent classification.
However, both for leaf differentiation and for insect counting (at the main leaf), the above-described deep learning techniques provide accuracy (i.e., the terms of false positives, false negatives).
For enhance understanding, the description has described a scenario with field user 192 operating mobile device 302. This however not required, image 412 could be taken otherwise, for example by aircraft flying over the field. The example of an unmanned aerial vehicle (UAV) is noted.
In such a scenario, the mobility to catch images on the field would be implemented by the UAV. User interface 392 (cf.
As explained above—for example, in connection with
In other words, the the relation of the physical size of the biological objects (132) to the physical size of the parts (122) is such that the representation of the biological objects on the part-images (422) are such that the representation is smaller than the tile dimension.
Persons of skill in the art can estimate the image resolution (i.e., the number of pixels per physical dimension).
There is also a limitation to the minimum. In the extreme case (minimum), the biological object of the smallest allowable size would be represented (in theory) by one pixel. More practical sizes have been explained above (cf.
From a high level perspective, the biological objects (132) are (or were) living organisms that are located on the parts (122) (of the plant), or the biological objects are traces by that organisms. (Optionally, the organism may be considered as no longer living, cf. the example with the pupa). More in detail, the biological objects (132) on the parts (122) (of the plant) are selected from the following: insects, arachnids, and mollusca. In an alternative, the biological objects are selected from spots or stripes on the surface of the plant parts. In that alternative, it does not matter if the objects are considered to be organisms or not, spots or stripes can be disease symptoms. For example, brown spots or brown stripes on a plant part indicates that the plant is potentially damaged. For example, some fungi cause brown stripes like yellow rust.
In view of the size (min/max) limitations, not all insects, arachnids or mollusca (or spots and stripes) would fit. For example, a large butterfly insect would potentially cover a single leaf (not countable), but small whiteflies (or a thrips) would be countable, as explained with much detail.
The description now shortly returns to the left side of
So far, the description has used two examples of plant/insect combinations (cf. the above section “Insects, plants and use cases”). In such embodiments with the special focus to estimate the numbers of insects (such as whitefly or thrips) on plants leaves, the segmentation is performed to segment out the main leaf from the rest of the image. The resulting image is the leaf-image.
However, the embodiments are based on a more general approach: It does not matter if the plant image 461 shows plant parts 122 such the stem, one or more branches, one or more leaves, flowers (usually with petals) or fruits (or other parts mentioned above). These parts—when represented by plant images 411/412—have borders that the CNN can recognize (if appropriately being training before). Looking at
Expert user 191 does not have to identify all parts in one image, but he/she can makes the annotations for parts of the plant should be segmented out by the CNN. For example, if a plant image shows a one or more leaves and a fruit, expert uses 191 could provide a first annotation for the leaf, and a second annotations for the fruit. The annotations would be used separately: in a CNN to identify the main leaf (as described in much detail, trained on the basis of the first annotations) and a CNN to identify the fruit (the skilled person can apply the existing description accordingly, trained on the basis of the second annotations).
In other words, computer 301 interacts with expert user 191 to obtain part-annotations (in general) and/or to obtain leaf-annotations (in particular). Thereby, expert user 191 conveys ground truth information to the images, not only regarding the presence or absence of a main leaf (by the leaf-annotations), but also to the parts in general (by part-annotations).
More in general, there is a computer-implemented method 602B for providing part-segmented images (showing parts of plants). Optionally, the method belongs to an overall process to estimate the number (NEST) of objects 132 on parts 122 of a plant 112, wherein method 602B identifies images for the parts.
In a method step, the computer uses first convolutional neural network 262 to process plant-image 412 to derive part-image 422 being a contiguous set of pixels that show a part 422-1 of particular plant 112 completely. First convolutional neural network 262 has been trained (before that method step) by a plurality of part-annotated plant-images 461, wherein the plant-images 411 are annotated to identify parts 421-1.
In one embodiment, there is provided a computer-implemented method for providing a segmentation of plant-images to part-segmented images. A computer program product—when loaded into a memory of a computer and being executed by at least one processor of the computer—performs the steps of this computer-implemented method. The same principle applies to a computer system for providing part-segmented images (showing parts of plants), with the system adapted to perform the method.
As mentioned above, for the leaf-annotation, it does not matter if the leaf shows insects or not. The same principle applies to the parts in general: if the parts (no the images) show biological objects or not does not matter. There is an assumption that the biological objects are located above the parts (e.g., the leaf or the fruit), not at the borders.
Computing device 900 includes a processor 902, memory 904, a storage device 906, a high-speed interface 908 connecting to memory 904 and high-speed expansion ports 910, and a low speed interface 912 connecting to low speed bus 914 and storage device 906. Each of the components 902, 904, 906, 908, 910, and 912, are interconnected using various busses, and may be mounted on a common motherboard or in other manners as appropriate. The processor 902 can process instructions for execution within the computing device 900, including instructions stored in the memory 904 or on the storage device 906 to display graphical information for a GUI on an external input/output device, such as display 916 coupled to high speed interface 908. In other implementations, multiple processors and/or multiple buses may be used, as appropriate, along with multiple memories and types of memory. Also, multiple computing devices 900 may be connected, with each device providing portions of the necessary operations (e.g., as a server bank, a group of blade servers, or a multi-processor system).
The memory 904 stores information within the computing device 900. In one implementation, the memory 904 is a volatile memory unit or units. In another implementation, the memory 904 is a non-volatile memory unit or units. The memory 904 may also be another form of computer-readable medium, such as a magnetic or optical disk.
The storage device 906 is capable of providing mass storage for the computing device 900. In one implementation, the storage device 906 may be or contain a computer-readable medium, such as a floppy disk device, a hard disk device, an optical disk device, or a tape device, a flash memory or other similar solid state memory device, or an array of devices, including devices in a storage area network or other configurations. A computer program product can be tangibly embodied in an information carrier. The computer program product may also contain instructions that, when executed, perform one or more methods, such as those described above. The information carrier is a computer- or machine-readable medium, such as the memory 904, the storage device 906, or memory on processor 902.
The high speed controller 908 manages bandwidth-intensive operations for the computing device 900, while the low speed controller 912 manages lower bandwidth-intensive operations. Such allocation of functions is exemplary only. In one implementation, the high-speed controller 908 is coupled to memory 904, display 916 (e.g., through a graphics processor or accelerator), and to high-speed expansion ports 910, which may accept various expansion cards (not shown). In the implementation, low-speed controller 912 is coupled to storage device 906 and low-speed expansion port 914. The low-speed expansion port, which may include various communication ports (e.g., USB, Bluetooth, Ethernet, wireless Ethernet) may be coupled to one or more input/output devices, such as a keyboard, a pointing device, a scanner, or a networking device such as a switch or router, e.g., through a network adapter.
The computing device 900 may be implemented in a number of different forms, as shown in the figure. For example, it may be implemented as a standard server 920, or multiple times in a group of such servers. It may also be implemented as part of a rack server system 924. In addition, it may be implemented in a personal computer such as a laptop computer 922. Alternatively, components from computing device 900 may be combined with other components in a mobile device (not shown), such as device 950. Each of such devices may contain one or more of computing device 900, 950, and an entire system may be made up of multiple computing devices 900, 950 communicating with each other.
Computing device 950 includes a processor 952, memory 964, an input/output device such as a display 954, a communication interface 966, and a transceiver 968, among other components. The device 950 may also be provided with a storage device, such as a microdrive or other device, to provide additional storage. Each of the components 950, 952, 964, 954, 966, and 968, are interconnected using various buses, and several of the components may be mounted on a common motherboard or in other manners as appropriate.
The processor 952 can execute instructions within the computing device 950, including instructions stored in the memory 964. The processor may be implemented as a chipset of chips that include separate and multiple analog and digital processors. The processor may provide, for example, for coordination of the other components of the device 950, such as control of user interfaces, applications run by device 950, and wireless communication by device 950.
Processor 952 may communicate with a user through control interface 958 and display interface 956 coupled to a display 954. The display 954 may be, for example, a TFT LCD (Thin-Film-Transistor Liquid Crystal Display) or an OLED (Organic Light Emitting Diode) display, or other appropriate display technology. The display interface 956 may comprise appropriate circuitry for driving the display 954 to present graphical and other information to a user. The control interface 958 may receive commands from a user and convert them for submission to the processor 952. In addition, an external interface 962 may be provide in communication with processor 952, so as to enable near area communication of device 950 with other devices. External interface 962 may provide, for example, for wired communication in some implementations, or for wireless communication in other implementations, and multiple interfaces may also be used.
The memory 964 stores information within the computing device 950. The memory 964 can be implemented as one or more of a computer-readable medium or media, a volatile memory unit or units, or a non-volatile memory unit or units. Expansion memory 984 may also be provided and connected to device 950 through expansion interface 982, which may include, for example, a SIMM (Single In Line Memory Module) card interface. Such expansion memory 984 may provide extra storage space for device 950, or may also store applications or other information for device 950. Specifically, expansion memory 984 may include instructions to carry out or supplement the processes described above, and may include secure information also. Thus, for example, expansion memory 984 may act as a security module for device 950, and may be programmed with instructions that permit secure use of device 950. In addition, secure applications may be provided via the SIMM cards, along with additional information, such as placing the identifying information on the SIMM card in a non-hackable manner.
The memory may include, for example, flash memory and/or NVRAM memory, as discussed below. In one implementation, a computer program product is tangibly embodied in an information carrier. The computer program product contains instructions that, when executed, perform one or more methods, such as those described above. The information carrier is a computer- or machine-readable medium, such as the memory 964, expansion memory 984, or memory on processor 952, that may be received, for example, over transceiver 968 or external interface 962.
Device 950 may communicate wirelessly through communication interface 966, which may include digital signal processing circuitry where necessary. Communication interface 966 may provide for communications under various modes or protocols, such as
GSM voice calls, SMS, EMS, or MMS messaging, CDMA, TDMA, PDC, WCDMA, CDMA2000, or GPRS, among others. Such communication may occur, for example, through radio-frequency transceiver 968. In addition, short-range communication may occur, such as using a Bluetooth, WiFi, or other such transceiver (not shown). In addition, GPS (Global Positioning System) receiver module 980 may provide additional navigation- and location-related wireless data to device 950, which may be used as appropriate by applications running on device 950.
Device 950 may also communicate audibly using audio codec 960, which may receive spoken information from a user and convert it to usable digital information. Audio codec 960 may likewise generate audible sound for a user, such as through a speaker, e.g., in a handset of device 950. Such sound may include sound from voice telephone calls, may include recorded sound (e.g., voice messages, music files, etc.) and may also include sound generated by applications operating on device 950.
The computing device 950 may be implemented in a number of different forms, as shown in the figure. For example, it may be implemented as a cellular telephone 980. It may also be implemented as part of a smart phone 982, personal digital assistant, or other similar mobile device.
Various implementations of the systems and techniques described here can be realized in digital electronic circuitry, integrated circuitry, specially designed ASICs (application specific integrated circuits), computer hardware, firmware, software, and/or combinations thereof. These various implementations can include implementation in one or more computer programs that are executable and/or interpretable on a programmable system including at least one programmable processor, which may be special or general purpose, coupled to receive data and instructions from, and to transmit data and instructions to, a storage system, at least one input device, and at least one output device.
These computer programs (also known as programs, software, software applications or code) include machine instructions for a programmable processor, and can be implemented in a high-level procedural and/or object-oriented programming language, and/or in assembly/machine language. As used herein, the terms “machine-readable medium” and “computer-readable medium” refer to any computer program product, apparatus and/or device (e.g., magnetic discs, optical disks, memory, Programmable Logic
Devices (PLDs)) used to provide machine instructions and/or data to a programmable processor, including a machine-readable medium that receives machine instructions as a machine-readable signal. The term “machine-readable signal” refers to any signal used to provide machine instructions and/or data to a programmable processor.
To provide for interaction with a user, the systems and techniques described here can be implemented on a computer having a display device (e.g., a CRT (cathode ray tube) or LCD (liquid crystal display) monitor) for displaying information to the user and a keyboard and a pointing device (e.g., a mouse or a trackball) by which the user can provide input to the computer. Other kinds of devices can be used to provide for interaction with a user as well; for example, feedback provided to the user can be any form of sensory feedback (e.g., visual feedback, auditory feedback, or tactile feedback); and input from the user can be received in any form, including acoustic, speech, or tactile input.
The systems and techniques described here can be implemented in a computing device that includes a back end component (e.g., as a data server), or that includes a middleware component (e.g., an application server), or that includes a front end component (e.g., a client computer having a graphical user interface or a Web browser through which a user can interact with an implementation of the systems and techniques described here), or any combination of such back end, middleware, or front end components. The components of the system can be interconnected by any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include a local area network (“LAN”), a wide area network (“WAN”), and the Internet.
The computing device can include clients and servers. A client and server are generally remote from each other and typically interact through a communication network. The relationship of client and server arises by virtue of computer programs running on the respective computers and having a client-server relationship to each other.
A number of embodiments have been described. Nevertheless, it will be understood that various modifications may be made without departing from the spirit and scope of the invention.
In addition, the logic flows depicted in the figures do not require the particular order shown, or sequential order, to achieve desirable results. In addition, other steps may be provided, or steps may be eliminated, from the described flows, and other components may be added to, or removed from, the described systems. Accordingly, other embodiments are within the scope of the following claims.
Number | Date | Country | Kind |
---|---|---|---|
19200657.5 | Sep 2019 | EP | regional |
Filing Document | Filing Date | Country | Kind |
---|---|---|---|
PCT/EP2020/077197 | 9/29/2020 | WO |