IP Library › Granted Patent US 11,480,931
Granted Patent B2
US 11,480,931 · App. 16/968,164 · Granted Oct 25, 2022

Learning device, learning method, and program therefor

Inventors: Takashi Fujii (Kyoto, JP); Yuki Ueyama (Kyoto, JP); Yasuaki Abe (Takatsuki, JP); Nobuyuki Sakatani (Otsu, JP); Kazuhiko Imatake (Osaka, JP)
Assignee: OMRON Corporation
G05B13/0265G05B13/024G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,480,931
App. No.
16/968,164
Granted
Oct 25, 2022
Kind
B2
Abstract

This learning device provides a learned model to an adjuster containing a learned model learned to output a predetermined compensation amount to a controller, in a control system including the controller outputting a command value obtained by compensating a target value based on a compensation amount and a control object controlled to process an object to be processed. The learning device includes: an evaluation part obtaining operation data including the target value, command value and control variable and evaluates the quality of the control variable; a learning part generating candidate compensation amounts based on the operation data, and learning, as teacher data, the generated candidate compensation amount and the specific parameter of the object, and generating a learned model; and a setting part providing the learned model to the adjuster if the evaluated quality is within an allowable range.

Claims (29)

1. A control system comprises:

a learning device;

a controller which outputs a command value obtained by compensating a target value based on a compensation amount; and

a control object controlled to perform a predetermined process on an object to be processed, wherein the command value output by the controller is input to the control object, and the control object outputs a control variable as a response to the command value,

wherein the learning device provides, to an adjuster including a learned model learned to output the compensation amount to the controller based on a specific parameter of the object to be processed, the learned model, and

wherein the learning device operates to:

obtain operation data including the target value, the command value, and the control variable to evaluate the quality of the control variable by determining whether the control variable is within a predetermined allowable range;

generate a candidate compensation amount based on the operation data, performs learning with the generated candidate compensation amount and the specific parameter of the object to be processed as teacher data, and generates a learned model; and

provide the learned model to the adjuster when the control variable evaluated is within the predetermined allowable range and when the command value obtained by compensating the target value based on the compensation amount output by the generated learned model is imparted to the control object.

2. The learning device according to claim 1 , wherein when the specific parameter of the object to be processed provided to the control object is equal to a parameter, whose evaluation has not been performed yet, of the generated learned model, wherein the learning device outputs a compensation amount output by the generated learned model to the controller and to evaluate the quality.

3. The learning device according to claim 2 , wherein the learning device performs learning again when the quality evaluated based on the compensation amount output by the generated learned model is out of the predetermined allowable range, and regenerates the learned model.

4. The learning device according to claim 1 , wherein the learning device generates the candidate compensation amount by data-driven control.

5. The learning device according to claim 4 , wherein the data-driven control is any one of virtual reference feedback tuning (VRFT), fictitious reference iterative tuning (FRIT), and estimated response iterative tuning (ERIT).

6. A learning method executed to operate a learning device in a control system, the learning method comprises:

outputting, by a controller, a command value obtained by compensating a target value based on a compensation amount; and

controlling a control object to perform a predetermined process on an object to be processed, wherein the command value output by the controller is input to the control object, and the control object outputs a control variable as a response to the command value,

wherein the learning device provides, to an adjuster including a learned model learned to output the compensation amount to the controller based on a specific parameter of the object to be processed, the learned model, and

wherein the learning method further comprises operating the learning device to perform:

obtaining operation data including the target value by determining whether the control variable is within a predetermined allowable range, the command value, and the control variable to evaluate the quality of the control variable;

generating a candidate compensation amount based on the operation data, performing learning with the generated candidate compensation amount and the specific parameter of the object to be processed as teacher data, and generating a learned model; and

providing the learned model to the adjuster when the control variable evaluated is within the predetermined allowable range and when the command value obtained by compensating the target value based on the compensation amount output by the generated learned model is imparted to the control object.

7. A non-transitory computer-readable recording medium storing a program which, when executed by a processor causes the processor to perform a learning method to operate a learning device in a control system, the learning method comprises:

outputting, by a controller, a command value obtained by compensating a target value based on a compensation amount; and

controlling a control object to perform a predetermined process on an object to be processed, wherein the command value output by the controller is input to the control object, and the control object outputs a control variable as a response to the command value,

wherein the learning device provides, to an adjuster including a learned model learned to output the compensation amount to the controller based on a specific parameter of the object to be processed, the learned model, and

wherein the learning method further comprises operating the learning device to perform:

obtaining operation data including the target value, the command value, and the control variable to evaluate the quality of the control variable by determining whether the control variable is within a predetermined allowable range;

generating a candidate compensation amount based on the operation data, performing learning with the generated candidate compensation amount and the specific parameter of the object to be processed as teacher data, and generating a learned model; and

providing the learned model to the adjuster when the control variable evaluated is within the predetermined allowable range and when the command value obtained by compensating the target value based on the compensation amount output by the generated learned model is imparted to the control object.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 17, 2020
From: FUJII, TAKASHI; UEYAMA, YUKI; ABE, YASUAKI; SAKATANI, NOBUYUKI; IMATAKE, KAZUHIKO
To: OMRON CORPORATION
Reel/Frame 053507/0460 →
Priority Claims (1)
JP JP2018-047865 · Mar 15, 2018 · national
Continuity (1)
Related Publication 20210041838A1 · Feb 11, 2021