You are on page 1of 2

Methodology

Object recognition is a general term to describe a collection of related computer vision


tasks that involve identifying objects in digital photographs.
we can distinguish between these three computer vision tasks:
 Image Classification: Predict the type or class of an object in an image.
 Input: An image with a single object, such as a photograph.
 Output: A class label (e.g. one or more integers that are mapped to class
labels).
 Object Localization: Locate the presence of objects in an image and indicate
their location with a bounding box.
 Input: An image with one or more objects, such as a photograph.
 Output: One or more bounding boxes (e.g. defined by a point, width, and
height).
 Object Detection: Locate the presence of objects with a bounding box and types
or classes of the located objects in an image.
 Input: An image with one or more objects, such as a photograph.
 Output: One or more bounding boxes (e.g. defined by a point, width, and
height), and a class label for each bounding box.
One further extension to this breakdown of computer vision tasks is object segmentation,
also called “object instance segmentation” or “semantic segmentation,” where instances
of recognized objects are indicated by highlighting the specific pixels of the object
instead of a coarse bounding box.

Source code:
 First we have to Copy the RetinaNet model file and the image
you want to detect to the folder that contains the python file.
 In the first 3 lines, we imported the ImageAI object detection class in the first
line, imported the python os class in the second line and defined a variable to
hold the path to the folder where our python file, RetinaNet model file and
images are in the third line.
 In the next 5 lines of code above, we defined our object detection class in the
first line, set the model type to RetinaNet in the second line, set the model path
to the path of our RetinaNet model in the third line, load the model into the
object detection class in the fourth line, then we called the detection function
and parsed in the input image path and the output image path in the fifth line.
 In the next 2 lines of code, we iterate over all the results returned by
the detector.detectObjectsFromImage function in the first line, then
print out the name and percentage probability of the model on each object
detected in the image in the second line.
 ImageAI supports many powerful customization of the object detection process.
One of it is the ability to extract the image of each object detected in the image. By
simply parsing the extra parameter extract_detected_objects=True into
the detectObjectsFromImage function as seen below, the object detection class
will create a folder for the image objects, extract each image, save each to the new
folder created and return an extra array that contains the path to each of the images.

You might also like