IP Library › Granted Patent US 12,462,160
Granted Patent B2
US 12,462,160 · App. 17/263,868 · Granted Nov 4, 2025

Performing distributed processing using layers of a neural network divided between a first device and a second device

Inventors: Ryo Takahashi (Tokyo, JP); Yukio Oobuchi (Kanagawa, JP)
Assignee: Sony Corporation
G06N3/086G06F18/2163G06F18/217G06N3/048
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,462,160
App. No.
17/263,868
Granted
Nov 4, 2025
Kind
B2
Abstract

An information processing method according to the present disclosure includes the steps of: evaluating, by a computer, a neural network having a structure held in a divided manner by a first device and a second device based on information on transfer of information between the first device and the second device in the neural network; and determining, by the computer, the structure of the neural network based on the evaluation of the neural network.

Claims (32)

1 . An information processing method executed by processing circuitry, the method comprising:

evaluating transfer of information between layers of a neural network performing distributed processing, wherein the layers of the neural network are divided between a first device and a second device, to determine an evaluation value;

wherein:

a third layer among the layers of the neural network that lies in a portion deeper than a second layer from which second information having a maximum size is output, the third layer outputs third information having a size smaller than that of first information output from an input layer of the neural network, the third layer is determined to be a transfer point at which the third information is to be transmitted from the first device to the second device, and

the neural network is evaluated based on the third information on the determined transfer point;

determining if the evaluation value meets a predetermined condition;

determining a structure of the neural network that meets the predetermined condition; and

executing a distributed process using the determined structure in which the layers of the neural network are divided between the first device and the second device.

2 . The information processing method according to claim 1 , wherein the neural network is evaluated based on a number of layers that lie in portions shallower than the transfer point and a total number of the layers constituting the neural network.

3 . The information processing method according to claim 1 , wherein the neural network is evaluated based on an indicator value that represents a recognition performance of the neural network.

4 . The information processing method according to claim 1 , wherein the neural network is evaluated based on a calculation amount in the neural network.

5 . The information processing method according to claim 1 , wherein the neural network is evaluated based on information on a performance of arithmetic processing of the first device.

6 . The information processing method according to claim 5 , wherein the neural network is evaluated based on a number of times of floating-point operations and a number of times of operations other than the floating-point operations in each layer of the neural network held in the first device.

7 . The information processing method according to claim 5 , wherein the neural network is evaluated based on a relation between a number of times of multiplication and a number of times of operations other than the multiplication performed in each layer of the neural network held in the first device.

8 . The information processing method according to claim 1 , wherein the neural network is evaluated based on values obtained by multiplying the information on transfer, an indicator value that represents a recognition performance of the neural network, a calculation amount in the neural network, and information on a performance of arithmetic processing of the first device by respective predetermined weight values.

9 . The information processing method according to claim 8 , wherein the weight values are determined based on configurations of the first device and the second device, a communication standard between the first device and the second device, and information on an environment in which the neural network is provided.

10 . An information processing device comprising:

processing circuitry configured to:

evaluate transfer of information between layers of a neural network performing distributed processing, wherein the layers of the neural network are divided between a first device and a second device, to determine an evaluation value;

wherein a third layer among the layers of the neural network that lies in a portion deeper than a second layer from which second information having a maximum size is output, the third layer outputs third information having a size smaller than that of first information output from an input layer of the neural network, the third layer is determined to be a transfer point at which the third information is to be transmitted from the first device to the second device, and

the neural network is evaluated based on the third information on the determined transfer point;

determine if the evaluation value meets a predetermined condition;

determine a structure of the neural network that meets the predetermined condition; and

execute a distributed process using the determined structure in which the layers of the neural network are divided between the first device and the second device.

11 . A non-transitory computer readable medium storing instructions that, when executed by processing circuitry, perform an information processing method comprising:

evaluating transfer of information between layers of a neural network performing distributed processing, wherein the layers of the neural network are divided between a first device and a second device, to determine an evaluation value;

wherein:

a third layer among the layers of the neural network that lies in a portion deeper than a second layer from which second information having a maximum size is output, the third layer outputs third information having a size smaller than that of first information output from an input layer of the neural network, the third layer is determined to be a transfer point at which the third information is to be transmitted from the first device to the second device, and

the neural network is evaluated based on the third information on the determined transfer point;

determining if the evaluation value meets a predetermined condition;

determining a structure of the neural network that meets the predetermined condition; and

executing a distributed process using the determined structure in which the layers of the neural network are divided between the first device and the second device.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 12, 2021
From: TAKAHASHI, RYO; OOBUCHI, YUKIO
To: SONY CORPORATION
Reel/Frame 056219/0495 →
Priority Claims (1)
JP 2018-147180 · Aug 3, 2018 · national
Continuity (1)
Related Publication 20210312295A1 · Oct 7, 2021
References Cited (19)
US 11205110B2 · Willson · 2021 [cited by examiner]
US 20160284400A1 · Yakopcic · 2016 [cited by examiner]
US 20170228639A1 · Hara · 2017 [cited by examiner]
US 20180307969A1 · Shibahara · 2018 [cited by examiner]
US 20190057308A1 · Cho · 2019 [cited by examiner]
US 20190073335A1 · Foley · 2019 [cited by examiner]
US 20190083335A1 · Zhang · 2019 [cited by examiner]
US 20190197359A1 · Haneda · 2019 [cited by examiner]
US 20190354852A1 · Sadowski · 2019 [cited by examiner]
CN 106960281A · 2017 [cited by applicant]
CN 108021983A · 2018 [cited by applicant]
WO WO2017134554A1 · 2017 [cited by applicant]
WO WO2017154284A · 2017 [cited by applicant]
WO WO2018003457A1 · 2018 [cited by examiner]
He Li, Learning IoT in Edge: Deep Learning for the Internet of Things with Edge Computing, Jan. 2018, p. 6 (Year: 2018). [cited by examiner]
International Written Opinion and English translation thereof mailed Sep. 10, 2019 in connection with International Application No. PCT/JP2019/027415. [cited by applicant]
International Preliminary Report on Patentability and English translation thereof mailed Feb. 18, 2021 in connection with International Application No. PCT/JP2019/027415. [cited by applicant]
International Search Report and English translation thereof mailed Sep. 10, 2019 in connection with International Application No. PCT/JP2019/027415. [cited by applicant]
Lane et al., Deepx: A software accelerator for low-power deep learning inference on mobile devices. 15th ACM/IEEE International Conference on Information Processing in Sensor Networks (IPSN). Apr. 11, 2016. 12 pages. [cited by applicant]