IP Library › Granted Patent US 12,663,802
Granted Patent B2
US 12,663,802 · App. 18/745,868 · Granted Jun 23, 2026

Clutter tidying robot utilizing floor segmentation for mapping and navigation system

Inventors: Mahaveer Suthar (Bangalore, IN); Akash Jadhav (Kingston, CA); Justin David Hamilton (Wellington, NZ); Shantanu Singh (Bangalore, IN); Dhruv Krishna (Bangalore, IN)
Assignee: Clutterbot, Inc.
G05D1/2464A47L11/4011A47L11/4036G01S17/86G01S17/89G05D1/245G05D1/2465G06V10/25G06V10/26G06V10/30G06V10/764G06V10/803G06V20/58G06V20/70A47L2201/04G05D2105/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,663,802
App. No.
18/745,868
Filed
Jun 17, 2024
Granted
Jun 23, 2026
Kind
B2
Art Unit
3669
USPC
701/25
Abstract

A method and apparatus are disclosed for a clutter tidying robot utilizing floor segmentation for its mapping and navigation system, whereby a perception module and navigation module transform lidar and image data from lidar sensors and cameras of a robot sensing system using segmentation and pseudo-laserscan or point cloud transformations to generate global and local maps. The robot pose and maps are transmitted to a robot brain that directs an action module to produce robot action commands controlling the operation of a clutter tidying robot using the pose and map data. In this manner multi-stage planning and sophisticated obstacle avoidance techniques may be incorporated into autonomous robot operations.

Claims (161)

1 . A method comprising:

receiving, at a perception module configured in operational logic of a robotic control system, image data from a robot's sensors,

wherein the perception module comprises a detection module, a scene segmentation module, and a mapping module,

wherein the robot's sensors include at least one of cameras, lidar sensors, inertial measurement unit (IMU) sensors, wheel encoders, and other sensors, and

wherein the operational logic of the robotic control system comprises executable instructions stored by a storage device of the robotic control system and executed by a processor of the robotic control system;

detecting, by the detection module, objects from the image data, as two-dimensional (2D) bounding boxes with object classes;

generating, by the detection module:

predicted three-dimensional (3D) object locations, using the 2D bounding boxes and a ground plane; and

two-dimensional-three-dimensional (2D-3D) bounding boxes with class labels, the 2D-3D bounding boxes based on the 2D bounding boxes and the predicted 3D object locations;

generating, by the scene segmentation module:

a multi-class segmentation map, using a segmentation model to segregate a floor boundary and other relevant regions in the image data;

an edge map including the floor boundary and other relevant boundaries, with semantic information; and

semantic boundary masks, from the multi-class segmentation map, wherein the semantic boundary masks identify relevant boundaries and their semantic information;

generating, by the mapping module, using the 2D-3D bounding boxes and the semantic boundary masks:

a scene layout map, wherein the scene layout map includes global elements relevant for global mapping; and

a local occupancy map, wherein the local occupancy map includes local elements useful for local path planning and local obstacle avoidance;

receiving, at a navigation module configured in the operational logic of the robotic control system, the scene layout map, and the local occupancy map, wherein the navigation module includes:

a simultaneous localization and mapping module (SLAM);

a global mapper module; and

a fusion and inflation module;

generating, by the SLAM, using lidar data, IMU data, and wheel encoding data:

a SLAM global map, which comprises a 2D occupancy grid representation of an environment with obstacle information at lidar height and real-time location information of the robot; and

a robot pose;

generating, by the global mapper module, using the SLAM global map, the lidar data, and the scene layout map:

a navigation global map, which represents an improved 2D occupancy grid representation of the environment over the SLAM global map;

generating, by the fusion and inflation module using the lidar data, the navigation global map, and the local occupancy map:

a fused local occupancy map, with the lidar data and information from the navigation global map and the local occupancy map, making the fused local occupancy map useful for obstacle avoidance;

an inflated global map, which includes buffer regions for the robot around obstacles; and

an inflated local map based on the fused local occupancy map, which includes the buffer regions for the robot around the obstacles;

receiving, at a robot brain configured in the operational logic of the robotic control system, the inflated global map, the robot pose, and the inflated local map;

generating, by the robot brain, robot action commands based on at least one of the inflated global map, the robot pose, and the inflated local map;

receiving, at an action module configured in the operational logic of the robotic control system, the robot action commands; and

controlling robot actuators in response to the robot action commands.

2 . The method of claim 1 , further comprising:

receiving, at the robot brain, an interface signal from a robot user interface; and

generating, by the robot brain, the robot action commands based on at least one of the inflated global map, the robot pose, the inflated local map, and the interface signal.

3 . The method of claim 1 , further comprising:

detection filtering, by the detection module, to remove the 2D-3D bounding boxes that are not on the ground or are inside of a shovel on the robot.

4 . The method of claim 1 , further comprising:

running an edge detection algorithm, by the scene segmentation module, on the multi-class segmentation map, resulting in the semantic boundary masks with the semantic information.

5 . The method of claim 1 , further comprising:

adding, using a labeling module included in the mapping module, additional semantic labels to the segmentation boundary masks by:

determining the points in the edge map that are inside a corresponding 2D bounding box or a corresponding 2D-3D bounding box; and

assigning all of the points inside of the corresponding 2D bounding box or the corresponding 2D-3D bounding box a same label as that of the corresponding 2D bounding box or the corresponding 2D-3D bounding box, thereby resulting in a semantically rich boundary map;

warping, using a top view transformation module included in the mapping module, the semantically rich boundary map into a point cloud with real-world coordinates and semantic label classes by using at least one of:

a lookup table that stores pixel mappings from an image space to real-world coordinates; and

a homography matrix that maps pixels from the image space to the real-world coordinates in real time;

filtering, by a scene layout module included in the mapping module, the point cloud with the real-world coordinates and the semantic label classes, keeping the semantically relevant points needed for the global mapping and discarding points not needed for the global mapping; and

filtering, by a local occupancy module included in the mapping module, the point cloud with real-world coordinates and semantic label classes, keeping the semantically relevant points needed for the local path planning and the obstacle avoidance.

6 . The method of claim 1 , further comprising:

processing, using a filter and fusion module included in the SLAM, the lidar data, the IMU data, and the wheel encoding data by:

removing noise and potentially unreliable data from the lidar data and the IMU data resulting in filtered lidar data and filtered IMU data;

removing the noise and potentially unreliable data from an angular velocity output of the wheel encoding data; and

fusing non-angular velocity output data of the wheel encoding data with the filtered IMU data to generate filtered and fused odometry data.

7 . The method of claim 6 , further comprising:

generating, using a main pipeline included in the SLAM, the 2D occupancy grid representation of the environment with the obstacle information at lidar height and the real-time location information of the robot by:

receiving the filtered lidar data and the filtered and fused odometry data;

creating a new 2D point registration for each new laser measurement at a given odometry reading;

estimating and correcting odometry slippages between each odometry reading by scan-to-scan matching the new 2D point registrations, thereby resulting in odometry slippage data points;

adding the odometry slippage data points to a pose-graph, resulting in an optimized pose-graph;

looking for loop closure in a chain of ‘N’ of the odometry slippage data points, wherein the loop closure represents a process of determining if a current location observed by the robot's sensors has been previously visited by the robot;

on condition the loop closure is detected:

correcting odometry poses for each new 2D point registration based on the optimized pose-graph, thereby resulting in loop closure pose corrections; and

forming a common 2D occupancy grid using the 2D point registrations and probabilistically updated 2D point registrations, wherein the common 2D occupancy grid is the 2D occupancy grid representation of the environment with the obstacle information at lidar height and the real-time location information of the robot.

8 . The method of claim 7 , further comprising:

processing, using a sensor data filter, the lidar data, and the scene layout map by:

removing the noise and potentially unreliable data from the lidar data; and

passing the scene layout map through a semantic filter that caters to the filtering of points from the scene layout map based on the semantic information provided, resulting in filtered lidar data and a filtered scene layout map where unreliable and irrelevant semantic labels have been removed during mapping.

9 . The method of claim 8 , further comprising:

generating, using a multi-sensor data registration, an enhanced 2D occupancy grid representation of the environment with the obstacle information at lidar height, from the filtered lidar data and the filtered scene layout map, by:

receiving the filtered lidar data and the filtered scene layout map;

creating a new enhanced 2D point registration for each new laser measurement and the filtered scene layout map using the real-time location information of the robot from the SLAM; and

updating all registrations from all of the robot's sensors, probabilistically, together in an enhanced common 2D occupancy grid based on predetermined confidence values of the robot's sensors.

10 . The method of claim 9 , further comprising:

receiving, by loop closure integration, the new enhanced 2D point registrations and the SLAM global map including the loop closure pose corrections;

reiterating, temporally, over the new 2D point registrations that are near the loop closure pose corrections for all of the updated registrations of the robot's sensors from the multi-sensor data registration to provide reiterated loop closure pose corrections;

updating map pose data of each of the robot's sensors using the reiterated loop closure pose corrections;

re-updating the updated registrations of each of the robot's sensors with the updated map pose data, thereby resulting in enhanced registrations; and

re-updating, probabilistically with the enhanced registrations, respective cells in the enhanced common 2D occupancy grid, thereby resulting in the navigation global map.

11 . An apparatus comprising:

a robot;

a processor; and

a memory storing instructions that, when executed by the processor, configure the apparatus to:

receive, at a perception module configured in operational logic of a robotic control system, image data from the robot's sensors,

wherein the perception module comprises a detection module, a scene segmentation module, and a mapping module,

wherein the robot's sensors include at least one of cameras, lidar sensor, inertial measurement unit (IMU) sensors, wheel encoder, and other sensors, and

wherein the operational logic of the robotic control system comprises executable instructions stored by a storage device of the robotic control system and executed by a processor of the robotic control system;

detect, by the detection module, objects from the image data, as two-dimensional (2D) bounding boxes with object classes;

generate, by the detection module:

predicted three-dimensional (3D) object locations, using the 2D bounding boxes and a ground plane; and

two-dimensional-three-dimensional (2D-3D) bounding boxes with class labels, the 2D-3D bounding boxes based on the 2D bounding boxes and the predicted 3D object locations;

generate, by the scene segmentation module:

a multi-class segmentation map, using a segmentation model to segregate a floor boundary and other relevant regions in the image data;

an edge map including the floor boundary and other relevant boundaries, with semantic information; and

semantic boundary masks, from the multi-class segmentation map, wherein the semantic boundary masks identify relevant boundaries and their semantic information;

generate, by the mapping module, using the 2D-3D bounding boxes and the semantic boundary masks:

a scene layout map, wherein the scene layout map includes global elements relevant for global mapping; and

a local occupancy map, wherein the local occupancy map includes local elements useful for local path planning and local obstacle avoidance;

receive, at a navigation module configured in the operational logic of the robotic control system, the scene layout map, and the local occupancy map, wherein the navigation module includes:

a simultaneous localization and mapping module (SLAM);

a global mapper module; and

a fusion and inflation module;

generate, by the SLAM, using lidar data, (IMU) data, and wheel encoding data:

a SLAM global map, which comprises a 2D occupancy grid representation of an environment with obstacle information at lidar height and real-time location information of the robot; and

a robot pose;

generate, by the global mapper module, using the SLAM global map, the lidar data, and the scene layout map:

a navigation global map, which represents an improved 2D occupancy grid representation of the environment over the SLAM global map;

generate, by the fusion and inflation module using the lidar data, the navigation global map, and the local occupancy map:

a fused local occupancy map, with the lidar data and information from the navigation global map and the local occupancy map, making the fused local occupancy map useful for obstacle avoidance;

an inflated global map, which includes buffer regions for the robot around obstacles; and

an inflated local map based on the fused local occupancy map, which includes the buffer regions for the robot around the obstacles;

receive, at a robot brain, the inflated global map, the robot pose, and the inflated local map;

generate, by the robot brain configured in the operational logic of the robotic control system, robot action commands based on at least one of the inflated global map, the robot pose, and the inflated local map;

receive, at an action module configured in the operational logic of the robotic control system, the robot action commands; and

control robot actuators in response to the robot action commands.

12 . The apparatus of claim 11 , wherein the instructions further configure the apparatus to:

receive, at the robot brain, an interface signal from a robot user interface; and

generate, by the robot brain, the robot action commands based on at least one of the inflated global map, the robot pose, the inflated local map, and the interface signal.

13 . The apparatus of claim 11 , wherein the instructions further configure the apparatus to:

detection filter, by the detection module, to remove the 2D-3D bounding boxes that are not on the ground or are inside of a shovel on the robot.

14 . The apparatus of claim 11 , wherein the instructions further configure the apparatus to:

run an edge detection algorithm, by the scene segmentation module, on the multi-class segmentation map, resulting in the semantic boundary masks with the semantic information.

15 . The apparatus of claim 11 , wherein the instructions further configure the apparatus to:

add, using a labeling module included in the mapping module, additional semantic labels to the segmentation boundary masks by:

determine the points in the edge map that are inside a corresponding 2D bounding box or a corresponding 2D-3D bounding box; and

assign all of the points inside of the corresponding 2D bounding box or the corresponding 2D-3D bounding box a same label as that of the corresponding 2D bounding box or the corresponding 2D-3D bounding box, thereby resulting in a semantically rich boundary map;

warp, using a top view transformation module included in the mapping module, the semantically rich boundary map into a point cloud with real-world coordinates and semantic label classes by using at least one of:

a lookup table that stores pixel mappings from an image space to real-world coordinates; and

a homography matrix that maps pixels from the image space to the real-world coordinates in real time;

filter, by a scene layout module included in the mapping module, the point cloud with the real-world coordinates and the semantic label classes, keeping the semantically relevant points needed for the global mapping and discarding points not needed for the global mapping; and

filter, by a local occupancy module included in the mapping module, the point cloud with real-world coordinates and semantic label classes, keeping the semantically relevant points needed for the local path planning and the obstacle avoidance.

16 . The apparatus of claim 11 , wherein the instructions further configure the apparatus to:

process, using a filter and fusion module included in the SLAM, the lidar data, the IMU data, and the wheel encoding data by:

remove noise and potentially unreliable data from the lidar data and the IMU data resulting in filtered lidar data and filtered IMU data;

remove the noise and potentially unreliable data from an angular velocity output of the wheel encoding data; and

fuse non-angular velocity output data of the wheel encoding data with the filtered IMU data to generate filtered and fused odometry data.

17 . The apparatus of claim 16 , wherein the instructions further configure the apparatus to:

generate, using a main pipeline included in the SLAM, the 2D occupancy grid representation of the environment with the obstacle information at lidar height and the real-time location information of the robot by:

receive the filtered lidar data and the filtered and fused odometry data;

create a new 2D point registration for each new laser measurement at a given odometry reading;

estimate and correct odometry slippages between each odometry reading by scan-to-scan matching the new 2D point registrations, thereby resulting in odometry slippage data points;

add the odometry slippage data points to a pose-graph, resulting in an optimized pose-graph;

look for loop closure in a chain of ‘N’ of the odometry slippage data points, wherein the loop closure represents a process of determining if a current location observed by the robot's sensors has been previously visited by the robot;

on condition the loop closure is detected:

correct the odometry poses for each new 2D point registration based on the optimized pose-graph, thereby resulting in loop closure pose corrections;

form a common 2D occupancy grid using the 2D point registrations and probabilistically updated 2D point registrations, wherein the common 2D occupancy grid is the 2D occupancy grid representation of the environment with the obstacle information at lidar height and the real-time location information of the robot.

18 . The apparatus of claim 17 , wherein the instructions further configure the apparatus to:

process, using a sensor data filter, the lidar data, and the scene layout map by:

remove the noise and potentially unreliable data from the lidar data; and

pass the scene layout map through a semantic filter that caters to the filtering of points from the scene layout map based on the semantic information provided, resulting in filtered lidar data and a filtered scene layout map where unreliable and irrelevant semantic labels have been removed during mapping.

19 . The apparatus of claim 18 , wherein the instructions further configure the apparatus to:

generate, using a multi-sensor data registration, an enhanced 2D occupancy grid representation of the environment with the obstacle information at lidar height, from the filtered lidar data and the filtered scene layout map by:

receive filtered lidar data and the filtered scene layout map;

create a new enhanced 2D point registration for each new laser measurement and the filtered scene layout map using the real-time location information of the robot from the SLAM; and

update all registrations from all of the robot's sensors, probabilistically, together in an enhanced common 2D occupancy grid based on predetermined confidence values of the robot's sensors.

20 . The apparatus of claim 19 , wherein the instructions further configure the apparatus to:

receive, by loop closure integration, the new enhanced 2D point registrations and the SLAM global map including the loop closure pose corrections;

reiterate, temporally, over the new 2D point registrations that are near the loop closure pose corrections for all of the updated registrations of the robot's sensors from the multi-sensor data registration to provide reiterated loop closure pose corrections;

update map pose data of each of the robot's sensors using the reiterated loop closure pose corrections;

re-update the updated registrations of each of the robot's sensors with the updated map pose data, thereby resulting in enhanced registrations; and

re-update, probabilistically with the enhanced registrations, respective cells in the enhanced common 2D occupancy grid, thereby resulting in the navigation global map.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 18, 2024
From: SUTHAR, MAHAVEER; JADHAV, AKASH; HAMILTON, JUSTIN DAVID; SINGH, SHANTANU; KRISHNA, DHRUV
To: CLUTTERBOT, INC.
Reel/Frame 067755/0787 →
Priority Claims (1)
IN 202341040880 · Jun 15, 2023 · national
Continuity (1)
Related Publication 20240419183A1 · Dec 19, 2024
References Cited (126)
US 6338013B1 · Ruffner · 2002 [cited by applicant]
US 6810305B2 · Kirkpatrick, Jr. · 2004 [cited by applicant]
US 9089969B1 · Theobald · 2015 [cited by applicant]
US 9827677B1 · Gilbertson et al. · 2017 [cited by applicant]
US 9827678B1 · Gilbertson et al. · 2017 [cited by applicant]
US 10207408B1 · Ebrahimi Afrouzi · 2019 [cited by applicant]
US 11020858B2 · Amacker et al. · 2021 [cited by applicant]
US 11407118B1 · Augenbraun et al. · 2022 [cited by applicant]
US 11409296B1 · Kim · 2022 [cited by examiner]
US 11720117B1 · Kim · 2023 [cited by examiner]
US 11789457B1 · Woo · 2023 [cited by examiner]
US 12064880B2 · Hamilton et al. · 2024 [cited by applicant]
US 12287639B1 · Queiroz et al. · 2025 [cited by applicant]
US 12310545B1 · Hamilton · 2025 [cited by applicant]
US 20020116089A1 · Kirkpatrick, Jr. · 2002 [cited by applicant]
US 20060116973A1 · Okamoto et al. · 2006 [cited by applicant]
US 20070239315A1 · Sato et al. · 2007 [cited by applicant]
US 20090137348A1 · Tsai · 2009 [cited by applicant]
US 20160260161A1 · Atchley et al. · 2016 [cited by applicant]
US 20170329333A1 · Passot et al. · 2017 [cited by applicant]
US 20180104815A1 · Yang · 2018 [cited by applicant]
US 20180284792A1 · Kleiner et al. · 2018 [cited by applicant]
US 20180312095A1 · Eletrabi · 2018 [cited by applicant]
US 20180329425A1 · Watts · 2018 [cited by applicant]
US 20190217476A1 · Jiang et al. · 2019 [cited by applicant]
US 20200026301A1 · Taylor · 2020 [cited by examiner]
US 20200086182A1 · New et al. · 2020 [cited by applicant]
US 20200156246A1 · Srivastav · 2020 [cited by applicant]
US 20200279436A1 · Bleyer · 2020 [cited by examiner]
US 20200376675A1 · Bai et al. · 2020 [cited by applicant]
US 20210038950A1 · Sullivan · 2021 [cited by applicant]
US 20210069904A1 · Duan et al. · 2021 [cited by applicant]
US 20210128989A1 · Legg et al. · 2021 [cited by applicant]
US 20210132213A1 · Advani · 2021 [cited by examiner]
US 20210138311A1 · Wang et al. · 2021 [cited by applicant]
US 20210339393A1 · Dan · 2021 [cited by applicant]
US 20210373560A1 · Koga et al. · 2021 [cited by applicant]
US 20220168893A1 · Hamilton et al. · 2022 [cited by applicant]
US 20220171907A1 · Chang et al. · 2022 [cited by applicant]
US 20220253056A1 · Ranjan et al. · 2022 [cited by applicant]
US 20220398775A1 · Streem · 2022 [cited by examiner]
US 20230015751A1 · Serackis et al. · 2023 [cited by applicant]
US 20230116896A1 · Hamilton et al. · 2023 [cited by applicant]
US 20230206755A1 · Jha · 2023 [cited by examiner]
US 20230401837A1 · Yan et al. · 2023 [cited by applicant]
US 20240142994A1 · Ebrahimi Afrouzi · 2024 [cited by applicant]
US 20240192690A1 · Ebrahimi Afrouzi · 2024 [cited by examiner]
US 20240218863A1 · Eletrabi et al. · 2024 [cited by applicant]
US 20240292990A1 · Hamilton · 2024 [cited by applicant]
US 20240424355A1 · Zhang et al. · 2024 [cited by applicant]
US 20240428560A1 · Buerkle et al. · 2024 [cited by applicant]
AU 2015218522A1 · 2015 [cited by applicant]
AU 2020101865A4 · 2020 [cited by applicant]
AU 2022360549A1 · 2024 [cited by applicant]
CN 202526850U · 2012 [cited by applicant]
CN 107137890A · 2017 [cited by applicant]
CN 107378958A · 2017 [cited by applicant]
CN 107650147A · 2018 [cited by applicant]
CN 108217183A · 2018 [cited by applicant]
CN 108786065A · 2018 [cited by applicant]
CN 110558900A · 2019 [cited by applicant]
CN 210616549U · 2020 [cited by applicant]
CN 111633627A · 2020 [cited by applicant]
CN 112828847A · 2021 [cited by applicant]
CN 114468859A · 2022 [cited by applicant]
CN 218552245U · 2023 [cited by applicant]
CN 116616642A · 2023 [cited by applicant]
CN 116709962A · 2023 [cited by applicant]
CN 117442132A · 2024 [cited by applicant]
DE 102017112740A1 · 2018 [cited by applicant]
EP 1942223B1 · 2008 [cited by applicant]
EP 3766591A1 · 2021 [cited by applicant]
JP 2019092999A · 2019 [cited by applicant]
JP 2020138046A · 2020 [cited by applicant]
JP 2023052271A · 2023 [cited by applicant]
KR 1020190130941 · 2019 [cited by applicant]
KR 20200126036A · 2020 [cited by applicant]
KR 20230139407A · 2023 [cited by applicant]
KR 20240133919A · 2024 [cited by applicant]
WO 2012103525A2 · 2012 [cited by applicant]
WO 2016105702A1 · 2016 [cited by applicant]
WO 2019126332A1 · 2019 [cited by applicant]
WO 2020071669A1 · 2020 [cited by applicant]
WO 2020190272A1 · 2020 [cited by applicant]
WO 2022029730A1 · 2022 [cited by applicant]
WO 2022115761A1 · 2022 [cited by applicant]
WO WO2023031620A1 · 2023 [cited by examiner]
WO 2023060285A1 · 2023 [cited by applicant]
WO 2023110103A1 · 2023 [cited by applicant]
WO 2025184405A1 · 2025 [cited by applicant]
Gonzalez et al., Multi-LiDAR Mapping for Scene Segmentation in Indoor Environments for Mobile Robots, May 2022, MDPI, pp. 1-19. (Year: 2022). [cited by examiner]
U.S. Appl. No. 63/119,533, filed Nov. 30, 2020, Justin David Hamilton. [cited by applicant]
Bahl, Laavanye et al. “Cu Bi: A Decluttering Robot”. Dec. 2018. Retrieved from the Internet (Year: 2018). [cited by applicant]
Bahl, Laavanye et al. “Cu Bi: Room De cluttering Robot”. May 2019. Retrieved from the Internet (Year: 2019). [cited by applicant]
U.S. Appl. No. 63/253,867, filed Oct. 8, 2021, Justin David Hamilton. [cited by applicant]
Bahl, Laavanye et al. “Cu Bi: Room De cluttering Robot”. May 2019. Retrieved from the Internet (Year: 2019) Carnegie Mellon University pp. 1-29. [cited by applicant]
PCT/US2022/077917 International Search Report dated Jan. 25, 2023. [cited by applicant]
PCT/US2022/077917 Written Opinion of the ISA. [cited by applicant]
Official Action Translation from Japanese Patent Office dated Apr. 2, 2024 for application JP 2023-533262 100107766. [cited by applicant]
PCT/US2021/061143 International Search Report Feb. 8, 2022. [cited by applicant]
PCT/US2021/061143 Written Opinion of the International Searching Authority Feb. 6, 2022. [cited by applicant]
China National Intellectual Property Administration Application No. 202180080246.2 filing date May 30, 2023 First Office Action, entire document dated Nov. 3, 2025. [cited by applicant]
China National Intellectual Property Administration Application No. 202180080246.2 filing date May 30, 2023 First Office Action, entire document dated Nov. 3, 2025 translation. [cited by applicant]
Changsheng Shen (Bobby), “CMU MRSD 2020: Team D—Cu Bi—A Decluttering Robot,” YouTube, Dec. 10, 2019. https://www.youtube.com/watch?v=IA-MaF _m8QA (accessed Jan. 19, 2026). (Year: 2019). [cited by applicant]
Canberk Alper et al: “Cloth Funnels: Canonicalized-Alignment for Multi-Purpose Garment Manipulation”, 2023 IEEE International Conference On Robotics and Automation (ICRA), IEEE, May 29, 2023 (May 29, 2023), pp. 5872-587… [cited by applicant]
CuBi: Room Decluttering Robot May 2019) (chromeextension://efaidnbmnnnibpcajpcglclefindmkaj/https://mrsdprojects.ri.cmu.edu/2018teamd/wpcontent/uploads/sites/34/2019/12/TeamD_FinalReport 10.pdf). [cited by applicant]
Guanqi Chen et al: “Language-Augmented Symbolic Planner for Open-world Task Planning”, Arxiv.Org, Cornell University Library, Jul. 13, 2024. [cited by applicant]
PCT/US25/43792, International Search Report, Dec. 10, 2025. [cited by applicant]
PCT/US25/43792, Written Opinion, Dec. 10, 2025. [cited by applicant]
PCT/US25/43802, International Search Report, Dec. 16, 2025. [cited by applicant]
PCT/US25/43802, Written Opinion, Dec. 16, 2025. [cited by applicant]
PCT/US25/43805, International Search Report, Dec. 16, 2025. [cited by applicant]
PCT/US25/43805, Written Opinion, Dec. 16, 2025. [cited by applicant]
PCT/US25/44087, International Search Report, Jan. 8, 2026. [cited by applicant]
PCT/US25/44087, Written Opinion, Jan. 8, 2026. [cited by applicant]
PCT/US25/44271, International Search Report, Jan. 5, 2026. [cited by applicant]
PCT/US25/44271, Written Opinion, Jan. 5, 2026. [cited by applicant]
PCT/US25/44310, International Search Report, Nov. 27, 2025. [cited by applicant]
PCT/US25/44310, Written Opinion, Nov. 27, 2025. [cited by applicant]
PCT/US25/44316, International Search Report, Jan. 8, 2026. [cited by applicant]
PCT/US25/44316, Written Opinion, Jan. 8, 2026. [cited by applicant]
PCT/US25/44322, International Search Report, Dec. 11, 2025. [cited by applicant]
PCT/US25/44322, Written Opinion, Dec. 11, 2025. [cited by applicant]
PCT/US25/44326, International Search Report, Dec. 15, 2025. [cited by applicant]
PCT/US25/44326, Written Opinion, Dec. 15, 2025. [cited by applicant]
Tsurumine Yoshihisa et al.: “Variationally Autoencoded Dynamic Policy Programming for Robotic Cloth Manipulation Planning based on Raw Images”, 2022 IEEE/SICE International Symposium On System Integration (SII), IEEE, J… [cited by applicant]