WO2015123647A1

Object ingestion through canonical shapes, systems and methods

Abstract

An object recognition ingestion system is presented. The object ingestion system captures image data of objects, possibly in an uncontrolled setting. The image data is analyzed to determine if one or more a priori know canonical shape objects match the object represented in the image data. The canonical shape object also includes one or more reference PoVs indicating perspectives from which to analyze objects having the corresponding shape. An object ingestion engine combines the canonical shape object along with the image data to create a model of the object. The engine generates a desirable set of model PoVs from the reference PoVs, and then generates recognition descriptors from each of the model PoVs. The descriptors, image data, model PoVs, or other contextually relevant information are combined into key frame bundles having sufficient information to allow other computing devices to recognize the object at a later time.

WO2015123647A1, drawing sheet 1
Sheet 1 of 3

Term

No projected expiry on record.

  1. Priority
  2. Filed
  3. Published
  4. Today

48 claims: 24 independent, 24 dependent

  1. 1
    CLAIMS What is claimed is:1. An object recognition ingestion system comprising: canonical shape database storing shape objects having geometrical attributes of canonical shapes, shape attributes, and having reference key frame points-of-view (PoVs);and an object ingestion engine coupled with the canonical shape database and programmed to perform the steps of: obtaining image data of at least one object;deriving a set of edges related to the at least one object from the image data;obtaining a shape result set from the canonical shape database where the shape result set include shape objects having shape attributes satisfying shape selection criteria determined as a function of geometrical information from the set of edges;selecting at least one target shape object from shape objects in the shape result set;generating an object model from the at least one target shape object and portions of the image data associated with the set of edges;deriving a set of model key frame PoVs from the object model and the reference key frame PoVs associated with the at least one target shape object;instantiating a descriptor object model from the object model, the descriptor model comprising recognition algorithm descriptors having locations on the object model relative to the model key frame PoVs;creating a set of key frames bundles from the descriptor object model as a function of the set of model key frame PoVs;and storing the set of key frame bundles in object recognition database.
  2. 25
    An object recognition ingestion system comprising:canonical shape database storing shape objects having geometrical attributes of canonical shapes, shape attributes, and having reference key frame points-of-view (PoVs);and an object ingestion engine coupled with the canonical shape database and programmed to perform the steps of: obtaining image data of at least one object;deriving a set of edges related to the at least one object from the image data;obtaining a shape result set from the canonical shape database where the shape result set include shape objects having shape attributes satisfying shape selection criteria determined as a function of geometrical information from the set of edges;selecting at least one target shape object from shape objects in the shape result set;generating an object model from the at least one target shape object and portions of the image data associated with the set of edges;deriving a set of model key frame PoVs from the object model and the reference key frame PoVs associated with the at least one target shape object;instantiating a descriptor object model from the object model, the descriptor model comprising recognition algorithm descriptors having locations on the object model relative to the model key frame PoVs;creating a set of key frames bundles from the descriptor object model as a function of the set of model key frame PoVs;and storing the set of key frame bundles in object recognition database.
  3. 27
    The system of any of claims 25 - 26, wherein the shape objects include geometrical primitives.
  4. 28
    The system of any of claims 25 - 27, wherein at least one of the shape objects comprise a compound shape object comprising at least two geometric primitives.
  5. 29
    The system of any of claims 25 - 28, wherein the geometrical primitives include at least one of the following:a line, a square, a cube, a circle, a sphere, a cylinder, a cone, a box, a torus, a platonic solid, a triangle, a pyramid, and a box.
  6. 30
    The system of any of claims 25 - 29, wherein at least some of the shape objects represent 3D objects.
  7. 31
    The system of any of claims 25 - 30, wherein the shape objects comprises topological classifications.
  8. 32
    The system of any of claims 25 - 31, wherein the shape objects comprise object templates representing object classes.
  9. 33
    The system of any of claims 25 - 32, wherein the object templates include at least one of the following:a vehicle, a building, an appliance, a plant, a toy, a face, a person, and an internal organ.
  10. 34
    The system of any of claims 25 - 33, wherein the reference key frame PoVs comprise a normal vector.
  11. 35
    The system of any of claims 25 - 34, wherein the reference key frame PoVs comprise key frame PoV generation rules.
  12. 36
    The system of any of claims 25 - 35, wherein the image data comprises at least one of the following types of data:visible data, video data, video frame data, still image data, acoustic imaging data, medical image data, and game imaging data.
  13. 37
    The system of any of claims 25 - 36, wherein the geometrical attributes include at least one of the following:a length, a width, a height, a thickness, a radius, a diameter, an angle, a hole, a center, a formula, a texture, a bounding box, a chirality, a periodicity, an orientation, a pitch, and a number of sides.
  14. 38
    The system of any of claims 25 - 37, further comprising a mobile device that includes the object ingestion engine.
  15. 39
    The system of any of claims 25 - 38, wherein the mobile device further includes the canonical shape database.
  16. 40
    The system of any of claims 25 - 39, wherein the mobile device further includes the object recognition database.
  17. 41
    The system of any of claims 25 - 40, wherein the recognition ingestion engine is further programmed to perform the step of obtaining the shape result set as a function of edge descriptors associated with the set of edges.
  18. 42
    The system of any of claims 25 - 41, wherein the canonical shape database indexes the shape objects based on edge descriptors.
  19. 43
    The system of any of claims 25 - 42, wherein the recognition ingestion engine is programmed to perform the step of selecting the at least one target shape object based on a user selection.
  20. 44
    The system of any of claims 25 - 43, wherein the recognition ingestion engine is programmed to perform the step of selecting the at least one target shape object based on a score.
  21. 45
    The system of any of claims 25 - 44, wherein the score is determined as a function of at least one of the following:a location, a time, and a descriptor match.
  22. 46
    The system of any of claims 25 - 45, wherein the recognition algorithm descriptors include at least one of the following type of descriptors:SIFT, FREAK, FAST, DAISY, and BRISK.
  23. 47
    The system of any of claims 25 - 46, wherein at least one key frame bundle within the set of key frame bundles includes the following:a normal vector, an image, and a descriptor.
  24. 48
    The system of any of claims 25 - 47, further comprising an imaging sensor programmed to perform the step of capturing the image data of the at least one object.
Independent claims24