IP Library Granted Patent US 12,639,915
Granted Patent B2
US 12,639,915 · App. 17/884,138 · Granted May 26, 2026

Systems and methods for object detection

Inventors: Ahmed Abouelela (Tokyo, JP); Hamdi Sahloul (Tokyo, JP); Jose Jeronimo Moreira Rodrigues (Tokyo, JP); Xutao Ye (Tokyo, JP); Jinze Yu (Tokyo, JP)
Assignee: MUJN, INC.
G06V10/7515B25J9/1664B25J9/1669B25J9/1697G06T1/0014G06T7/0002G06T7/70G06T7/74G06V10/443G06V10/751G06V10/757G06V10/7715G06V10/776G06V20/50G06V20/653G06T2207/30168G06V2201/06
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,639,915
App. No.
17/884,138
Granted
May 26, 2026
Kind
B2
Abstract

A computing system including a processing circuit in communication with a camera having a field of view. The processing circuit is configured to perform operations related to detecting, identifying, and retrieving objects disposed amongst a plurality of objects. The processing circuit may be configured to perform operations related to object recognition template generation, feature generation, hypothesis generation, hypothesis refinement, and hypothesis validation.

Claims (48)

1 . A computing system comprising:

at least one processing circuit in communication with a robot, having an arm and an end-effector connected thereto, and a camera having a field of view and configured, when one or more objects are or have been in the field of view, to execute instructions stored on a non-transitory computer-readable medium for:

obtaining object image information of an object in a scene;

obtaining a detection hypothesis including a corresponding object recognition template representing a template object;

identifying a discrepancy between the template object and the object image information;

identifying a set of template locations in the template object corresponding to a set of object locations of the object image information;

adjusting the set of template locations to converge to the set of object locations by identifying respective vectors extending between the set of template locations and corresponding ones of the set of object locations;

iteratively adjusting the set of template locations according to magnitudes and directions of the respective vectors until the respective vectors cancel each other out and no further adjustments to the template locations can be made; and

generating an adjusted detection hypothesis including an adjusted corresponding object recognition template according to the set of template locations after adjustment.

2 . The computing system of claim 1 , further comprising:

identifying new respective vectors according to the adjusted set of template locations after the respective vectors cancel each other out and no further adjustments to the template locations can be made; and

determining a quality of alignment.

3 . The computing system of claim 2 , wherein the quality of alignment is determined based on a level of misalignment defined by magnitudes and directions of the new respective vectors.

4 . The computing system of claim 2 , wherein the quality of alignment is determined based on distance measurements between the adjusted set of template locations and the set of object locations.

5 . The computing system of claim 4 , wherein the distance measurements include Euclidean distance measurements.

6 . The computing system of claim 4 , wherein the distance measurements include cosine distances between surface normal vectors associated with the adjusted set of template locations and the set of object locations.

7 . The computing system of claim 6 , wherein the cosine distances indicate angles between the surface normal vectors, and wherein sizes of the angles correlate with the quality of alignment.

8 . The computing system of claim 4 , wherein the distance measurements are measurements from a first location of the adjusted set of template locations to a plane of a second location of the set of object locations.

9 . The computing system of claim 2 , wherein the quality of alignment is determined by a rate of convergence between the adjusted set of template locations and the set of object locations.

10 . The computing system of claim 1 further including:

obtaining the detection hypothesis by overlaying the object recognition template with image information of the scene to identify the object image information based on comparisons between template gradient information and template surface normal vector information of the object recognition template and object gradient information and object surface normal vector information extracted from the image information.

11 . A method comprising:

obtaining object image information of an object in a scene;

obtaining a detection hypothesis including a corresponding object recognition template representing a template object;

identifying a discrepancy between the template object and the object image information;

identifying a set of template locations in the template object corresponding to a set of object locations of the object image information;

adjusting the set of template locations to converge to the set of object locations by identifying respective vectors extending between the set of template locations and corresponding ones of the set of object locations;

iteratively adjusting the set of template locations according to magnitudes and directions of the respective vectors until the respective vectors cancel each other out and no further adjustments to the template locations can be made; and

generating an adjusted detection hypothesis including an adjusted corresponding object recognition template according to the set of template locations after adjustment.

12 . The method of claim 11 , further comprising:

identifying new respective vectors according to the adjusted set of template locations after the respective vectors cancel each other out and no further adjustments to the template locations can be made; and

determining a quality of alignment.

13 . The method of claim 12 , further including:

determining the quality of alignment based on a level of misalignment defined by magnitudes and directions of the new respective vectors.

14 . The method of claim 12 , further including:

determining the quality of alignment based on distance measurements between the adjusted set of template locations and the set of object locations.

15 . The method of claim 12 , further including:

determining the quality of alignment by a rate of convergence between the adjusted set of template locations and the set of object locations.

16 . The method of claim 11 wherein obtaining the detection hypothesis further includes:

overlaying the object recognition template with image information of the scene to identify the object image information based on comparisons between template gradient information and template surface normal vector information of the object recognition template and object gradient information and object surface normal vector information extracted from the image information.

17 . A non-transitory computer readable medium, configured with executable instructions for implementing a method for refining a detection hypothesis, operable by at least one processing circuit via a communication interface configured to communicate with a robotic system, the method comprising:

receiving object image information of an object in a scene;

receiving a detection hypothesis including a corresponding object recognition template representing a template object;

performing an operation to identify a discrepancy between the template object and the object image information;

performing an operation to identify a set of template locations in the template object corresponding to a set of object locations of the object image information;

performing an operation to adjust the set of template locations to converge to the set of object locations by identifying respective vectors extending between the set of template locations and corresponding ones of the set of object locations;

performing an operation to iteratively adjust the set of template locations according to magnitudes and directions of the respective vectors until the respective vectors cancel each other out and no further adjustments to the template locations can be made; and

outputting to the robotic system an adjusted detection hypothesis including an adjusted corresponding object recognition template according to the set of template locations after adjustment.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 22, 2025
From: ABOUELELA, AHMED; SAHLOUL, HAMDI; MOREIRA RODRIGUES, JOSE JERONIMO; YE, XUTAO; YU, JINZE
To: MUJIN, INC.
Reel/Frame 070912/0959 →
Continuity (2)
Provisional Application 63230931 · Aug 9, 2021
Related Publication 20230044420A1 · Feb 9, 2023
References Cited (14)
US 11006039B1 · Yu et al. · 2021 [cited by applicant]
US 20040225472A1 · Kraft · 2004 [cited by examiner]
US 20110157178A1 · Tuzel · 2011 [cited by examiner]
US 20120098961A1 · Handa et al. · 2012 [cited by applicant]
US 20170220887A1 · Fathi et al. · 2017 [cited by applicant]
US 20190272411A1 · Zou · 2019 [cited by examiner]
Hodaň, Tomáš, et al. “Detection and fine 3D pose estimation of texture-less objects in RGB-D images.” 2015 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). IEEE, 2015. (Year: 2015). [cited by examiner]
Poli, Riccardo, James Kennedy, and Tim Blackwell. “Particle swarm optimization: An overview.” Swarm intelligence 1 (2007): 33-57. (Year: 2007). [cited by examiner]
Cao, Thanh-Tung, et al. “Parallel banding algorithm to compute exact distance transform with the GPU.” Proceedings of the 2010 ACM SIGGRAPH symposium on Interactive 3D Graphics and Games. 2010. (Year: 2010). [cited by examiner]
Zielinski, Karin, and Rainer Laur. “Stopping criteria for a constrained single-objective particle swarm optimization algorithm.” Informatica 31.1 (2007). (Year: 2007). [cited by examiner]
He, Ruotao, Juan Rojas, and Yisheng Guan. “A 3D Object Detection and Pose Estimation Pipeline Using RGB-D Images.” arXiv preprint arXiv:1703.03940v1 (2017). (Year: 2017). [cited by examiner]
Hinterstoisser, Stefan, et al. “Gradient response maps for real-time detection of textureless objects.” IEEE transactions on pattern analysis and machine intelligence 34.5 (2011): 876-888. (Year: 2011). [cited by examiner]
Besl, Paul J., and Neil D. McKay. “Method for registration of 3-D shapes.” Sensor fusion IV: control paradigms and data structures. vol. 1611. Spie, 1992. (Year: 1992). [cited by examiner]
Gelfand, Natasha, et al. “Geometrically stable sampling for the ICP algorithm.” Fourth International Conference on 3-D Digital Imaging and Modeling, 2003. 3DIM 2003. Proceedings.. IEEE, 2003. (Year: 2003). [cited by examiner]