IP Library › Granted Patent US 12,311,555
Granted Patent B2
US 12,311,555 · App. 17/969,100 · Granted May 27, 2025

Robotic navigation and transport of objects

Inventors: Sriram Nochur Narayanan (San Jose, CA); Ramin Moslemi (Pleasanton, CA); Junha Roh (San Francisco, CA)
Assignee: NEC Corporation
B25J9/1666
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,311,555
App. No.
17/969,100
Filed
Oct 19, 2022
Granted
May 27, 2025
Kind
B2
Art Unit
3656
USPC
700/255
Abstract

Navigational systems and methods include building a topological graph of an environment using nodes that represent locations in the space and associated directions, with frontiers associated with particular nodes and directions within the topological graph. An action is determined using a policy trained with an action reward function that weighs exploration to find new objects and moving objects to a goal. An agent navigates within the environment in accordance with the determined action.

Claims (102)

1. A navigation method, comprising:

building a topological graph of an environment using nodes that represent locations in the space and associated directions, with frontiers associated with particular nodes and directions within the topological graph;

determining an action using a policy trained with an action reward function that weighs exploration to find new objects and moving objects to a goal, including calculating an exploration score that represents a likelihood of finding an object upon exploring a frontier and predicting frontier scores on a per-frame basis using a convolutional neural network; and

navigating an agent within the environment in accordance with the determined action.

2. The method of claim 1 , wherein calculating the exploration score includes calculating a respective exploration score for each node-direction combination in the topological graph.

3. The method of claim 1 , wherein calculating the exploration score includes processing the topological graph using a graph convolutional neural network.

4. The method of claim 1 , wherein determining the action selects an action from the group consisting of explore, pickup, and drop and wherein the pickup and drop actions interact with a target object to move the target object from an origin location to a goal location.

5. The method of claim 1 , wherein the agent is an autonomous device and navigating the agent causes the agent to move within the environment.

6. The method of claim 1 , wherein the action reward function is expressed as:

r

t

e

=

success

·

r

s

⁢

u

⁢

c

⁢

c

⁢

e

⁢

s

⁢

s

e

+

r

slack

e

+

∑

o

found

o

·

r

found

e

where found o found is an indicator if an object o was found at a timestep t, r found e is a reward for finding a new object, r stack e is a time penalty for every step that encourages finding objects faster, success is an indicator if all objects were found, and r success e is an associated success bonus.

7. The method of claim 1 , wherein navigating includes using a navigation policy trained with a navigation reward function that weighs distances to objects.

8. The method of claim 7 , wherein the navigation reward function is expressed as:

r t n = [reached-obj] ·r obj n +r slack n +r d2o n +r collision n

where [reached-obj] is an indicator that a target object has been reached, r obj n is the success reward if the agent reaches closer than a threshold distance d th from the target object, r slack n is a constant time penalty for each step, r d2o n =(d t-1 −d t ) is the decrease in geodesic distance with the target object and r collision n is a penalty for collision with an object in the environment.

9. The method of claim 1 , further comprising using the exploration score to identify a frontier sub-goal.

10. The method of claim 1 , further comprising using an object closeness score to indicate a distance to an object within the agent's field of view.

11. A navigation system, comprising:

a hardware processor; and

a memory to store a computer program that, when executed by the hardware processor, causes the hardware processor to:

build a topological graph of an environment using nodes that represent locations in the space and associated directions, with frontiers associated with particular nodes and directions within the topological graph;

determine an action using a policy trained with an action reward function that weighs exploration to find new objects and moving objects to a goal and calculates an exploration score that represents a likelihood of finding an object upon exploring a frontier using a convolutional neural network to predict frontier scores on a per-frame basis; and

navigate an agent within the environment in accordance with the determined action.

12. The system of claim 11 , wherein the calculation of the exploration score includes a respective exploration score for each node-direction combination in the topological graph.

13. The system of claim 11 , wherein the calculation of the exploration score includes a graph convolutional neural network to process the topological graph.

14. The system of claim 11 , wherein the computer program further causes the hardware processor to select an action from the group consisting of explore, pickup, and drop, and wherein the pickup and drop actions interact with a target object to move the target object from an origin location to a goal location.

15. The system of claim 11 , wherein the agent is an autonomous device and navigating the agent causes the agent to move within the environment.

16. The system of claim 11 , wherein the action reward function is expressed as:

r

t

e

=

success

·

r

s

⁢

u

⁢

c

⁢

c

⁢

e

⁢

s

⁢

s

e

+

r

slack

e

+

∑

o

found

o

·

r

found

e

where found o is an indicator if an object o was found at a timestep t, r found e is a reward for finding a new object, r stack e is a time penalty for every step that encourages finding objects faster, success is an indicator if all objects were found, and r success e is an associated success bonus.

17. The system of claim 11 , wherein navigating includes using a navigation policy trained with a navigation reward function that weighs distances to objects.

18. The system of claim 17 , wherein the navigation reward function is expressed as:

r t n = [reached-obj] ·r obj n +r slack n +r d2o n +r collision n

where [reached-obj] is an indicator that a target object has been reached, r obj n is the success reward if the agent reaches closer than a threshold distance d th from the target object, r slack n is a constant time penalty for each step, r d2o n =(d t-1 −d t ) is the decrease in geodesic distance with the target object and r collision n is a penalty for collision with an object in the environment.

19. The system of claim 11 , wherein the exploration score is used to identify a frontier sub-goal.

20. The system of claim 11 , wherein an object closeness score is used to indicate a distance to an object within the agent's field of view.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 25, 2025
From: NEC LABORATORIES AMERICA, INC.
To: NEC CORPORATION
Reel/Frame 070945/0331 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 19, 2022
From: NARAYANAN, SRIRAM NOCHUR; MOSLEMI, RAMIN; ROH, JUNHA
To: NEC LABORATORIES AMERICA, INC.
Reel/Frame 061468/0644 →
Continuity (3)
Provisional Application 63270615 · Oct 22, 2021
Provisional Application 63279328 · Nov 15, 2021
Related Publication 20230132280A1 · Apr 27, 2023
References Cited (60)
US 11199853B1 · Afrouzi · 2021 [cited by examiner]
US 20110035051A1 · Kim · 2011 [cited by examiner]
US 20190025851A1 · Ebrahimi Afrouzi · 2019 [cited by applicant]
US 20190120633A1 · Afrouzi · 2019 [cited by examiner]
US 20190133396A1 · Lim · 2019 [cited by examiner]
US 20190133397A1 · Choe · 2019 [cited by examiner]
US 20200114924A1 · Chen · 2020 [cited by examiner]
US 20210004012A1 · Marchetti-Bowick · 2021 [cited by examiner]
US 20220026920A1 · Ebrahimi Afrouzi · 2022 [cited by examiner]
US 20230132280A1 · Narayanan · 2023 [cited by examiner]
CN 110632931A · 2019 [cited by examiner]
JP 2005032196A · 2005 [cited by examiner]
WO WO2019219969A1 · 2019 [cited by examiner]
WO WO2021080507A1 · 2021 [cited by examiner]
JP-2005032196-A translation (Year: 2005). [cited by examiner]
CN-110632931-A translation (Year: 2019). [cited by examiner]
Improving_object_detection_with_deep_convolutional_networks (Year: 2015). [cited by examiner]
Anderson et al., “On Evaluation of Embodied Navigation Agents”, arXiv:1807.06757v1 [cs.AI] Jul. 18, 2018, pp. 1-7. [cited by applicant]
Anderson et al., “Vision-and-Language Navigation: Interpreting visually-grounded navigation instructions in real environments”, in Proceedings of the IEEE conference on computer vision and pattern recognition 2018, Jun.… [cited by applicant]
Bacon et al., “The Option-Critic Architecture”, in Proceedings of the AAAI Conference on Artificial Intelligence Feb. 13, 2017, vol. 31, No. 1, pp. 1-9. [cited by applicant]
Batra et al., “ObjectNav Revisited: On Evaluation of Embodied Agents Navigating to Objects”, arXiv:2006.13171v2 [cs.CV] Aug. 30, 2020, pp. 1-9. [cited by applicant]
Chang et al., “Matterport3D: Learning from RGB-D Data in Indoor Environments”, arXiv:1709.06158v1 [cs.CV] Sep. 18, 2017, pp. 1-25. [cited by applicant]
Chaplot et al., “Object Goal Navigation using Goal-Oriented Semantic Exploration”, arXiv:2007.00643v2 [cs.CV] Jul. 2, 2020, pp. 1-11. [cited by applicant]
Chaplot et al., “Neural Topological SLAM for Visual Navigation”, in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2020, Jun. 2020, p. 12875-12884. [cited by applicant]
Chen et al., “SoundSpaces: Audio-Visual Navigation in 3D Environments”, arXiv:1912.11474v3 [cs.CV] Aug. 21, 2020, pp. 1-27. [cited by applicant]
Das et al., “Embodied Question Answering”, arXiv:1711.11543v2 [cs.CV] Dec. 1, 2017, pp. 1-20. [cited by applicant]
Das et al., “Neural Modular Control for Embodied Question Answering”, in Conference on Robot Learning Oct. 23, 2018, pp. 53-62. [cited by applicant]
Deitke et al., “RoboTHOR: An Open Simulation-to-Real Embodied AI Platform”, arXiv:2004.06799v1 [cs.CV] Apr. 14, 2020, pp. 1-11. [cited by applicant]
Fang et al., “Scene Memory Transformer for Embodied Agents in Long-Horizon Tasks”, in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2019, Jun. 2019, pp. 538-547. [cited by applicant]
Gan et al., “The threedworld transport challenge: A visually guided task-and-motion planning benchmark for physically realistic embodied ai”, arXiv:2103.14025v1 [cs.CV] Mar. 25, 2021, pp. 1-12. [cited by applicant]
Garrett et al., “Integrated task and motion planning”, arXiv:2010.01083v1 [cs.RO] Oct. 2, 2020, pp. 1-30. [cited by applicant]
Garrett et al., “Ffrob: Leveraging symbolic planning for efficient task and motion planning”, The International Journal of Robotics Research, Nov. 2017, pp. 104-136. [cited by applicant]
Gupta et al., “Cognitive mapping and planning for visual navigation”, in Proceedings of the IEEE conference on computer vision and pattern recognition 2017, Jul. 2017, pp. 2616-2625. [cited by applicant]
Jain et al., “A cordial sync: Going beyond marginal policies for multi-agent embodied tasks”, in European Conference on Computer Vision Aug. 23, 2020, pp. 471-490. [cited by applicant]
Kipf et al., “Semi-supervised classification with graph convolutional networks”, in International Conference on Learning Representations (ICLR), Apr. 2017, pp. 1-22. [cited by applicant]
Kolve et al., “AI2-THOR: An Interactive 3D Environment for Visual AI”, arXiv:1712.05474v4 [cs.CV] Aug. 26, 2022, pp. 1-12. [cited by applicant]
Krantz et al., “Waypoint Models for Instruction-guided Navigation in Continuous Environments”, in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), Oct. 2021, pp. 1-10. [cited by applicant]
LaValle, “Planning Algorithms”, Cambridge university press, May 29, 2006, pp. 1-1007. [cited by applicant]
Marza et al., “Teaching agents how to map: Spatial reasoning for multi-object navigation”, arXiv:2107.06011v3 [cs.CV] May 27, 2022, pp. 1-8. [cited by applicant]
McGovern et al., “Automatic discovery of subgoals in reinforcement learning using diverse density”, in Proceedings of the Eighteenth International Conference on Machine Learning, Jun. 2001, p. 361-368. [cited by applicant]
Oh et al., “Control of memory, active perception, and action in minecraft”, in International conference on machine learning, Jun. 11, 2016, pp. 2790-2799. [cited by applicant]
Parr et al., “Reinforcement learning with hierarchies of machines”, in Advances in Neural Information Processing Systems, vol. 10, MIT Press, Jul. 1998, pp. 1-7. [cited by applicant]
Ramakrishnan et al., “Occupancy anticipation for efficient exploration and navigation”, in European Conference on Computer Vision Aug. 23, 2020, pp. 400-418. [cited by applicant]
Savva et al., “Habitat: A platform for embodied ai research”, in Proceedings of the IEEE/CVF International Conference on Computer Vision, (ICCV), Oct. 2019, pp. 1-9. [cited by applicant]
Schulman et al., “Proximal policy optimization algorithms”, arXiv:1707.06347v2 [cs.LG] Aug. 28, 2017. pp 1-12. [cited by applicant]
Shen et al., “iGibson 1.0: A Simulation Environment for Interactive Tasks in Large Realistic Scenes”, arXiv:2012.02924v6 [cs.AI] Aug. 10, 2021, pp. 1-11. [cited by applicant]
Shridhar et al., “Alfred: A benchmark for interpreting grounded instructions for everyday tasks”, arXiv:1912.01734v2 [cs.CV] Mar. 31, 2020, pp. 1-18. [cited by applicant]
Sutton et al., “Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning”, Artificial Intelligence 112, Dec. 1998, pp. 181-211. [cited by applicant]
Szot et al., “Habitat 2.0: Training home assistants to rearrange their habitat”, Advances in Neural Information Processing Systems, Dec. 6, 2021, pp. 1-16. [cited by applicant]
Wani et al., “Multion: Benchmarking semantic map memory using multi-object navigation” Advances in Neural Information Processing Systems 33, Dec. 2020, pp. 1-13. [cited by applicant]
Weihs et al., “Visual room rearrangement”, in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Jun. 2021, pp. 5922-5931. [cited by applicant]
Wijmans et al., “Dd-ppo: Learning near-perfect pointgoal navigators from 2.5 billion frames”, arXiv:1911.00357v2 [cs.CV] Jan. 20, 2020, pp. 1-21. [cited by applicant]
Wortsman et al., “Learning to learn how to learn: Self-adaptive visual navigation using meta-learning”, in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, Jun. 2019, pp. 6750-6759. [cited by applicant]
Xia et al., “Relmogen: Leveraging motion generation in reinforcement learning for mobile manipulation” arXiv:2008.07792v2 [cs.AI], Mar. 26, 2021, pp. 1-16. [cited by applicant]
Xia et al., “Gibson Env: realworld perception for embodied agents” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Jun. 2018, pp. 9068-9079. [cited by applicant]
Yamada et al., “Motion planner augmented reinforcement learning for robot manipulation in obstructed environments”, arXiv:2010.11940v1 [cs.RO] Oct. 22, 2020, pp. 1-15. [cited by applicant]
Yamauchi, “A frontier-based approach for autonomous exploration”, in Proceedings 1997 IEEE International Symposium on Computational Intelligence in Robotics and Automation CIRA'97.‘Towards New Computational Principles f… [cited by applicant]
Yang et al., “Visual semantic navigation using scene priors”, arXiv: 1810.06543v1 [cs.CV] Oct. 15, 2018, pp. 1-14. [cited by applicant]
Hart et al., “A Formal Basis for the Heuristic Determination of Minimum Cost Paths”, IEEE Transactions of SYstems Science and Cybernetics, vol. 4, No. 2, Jul. 1968, pp. 1-8. [cited by applicant]
Siciliano et al., “Springer Handbook of Robotics”, Springer-Verlag, Berlin, Heidelberg, Aug. 2008, pp. 1-70. [cited by applicant]