IP Library › Granted Patent US 12,596,962
Granted Patent B2
US 12,596,962 · App. 17/760,232 · Granted Apr 7, 2026

Data transmission using data prioritization

Inventors: Mats Folkesson (Täby, SE); Bin Sun (Luleå, SE); Rafia Inam (Västerås, SE); Elena Fersman (Stockholm, SE); Daniel Lindström (Luleå, SE)
Assignee: TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)
G06N20/20G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,596,962
App. No.
17/760,232
Granted
Apr 7, 2026
Kind
B2
Abstract

A method for providing data to a destination node connected to one or more source nodes via one or more networks. In one aspect, the method is performed by a feature identifier for selecting at least one feature from a set of features. The set of features comprises a first feature and a second feature. The method includes, for each feature included in the set of features, obtaining a value indicating a cost of providing the data set for the feature from a source node that stores the data set to a destination node that is connected to the source node via a network. The method also includes, based on the obtained values, selecting a subset of the set of features. The method further includes for each selected feature, initiating the transmission of the respective data set for the respective selected feature from a source node to the destination node. The destination node may execute a machine learning process that is operable to use the respective data sets to produce a model.

Claims (53)

1 . A method for providing data to a destination node connected to one or more source nodes via one or more networks, the method being performed by a feature identifier for selecting at least one feature from a set of features, the set of features comprising a first feature and a second feature, wherein

each feature included in the set of features is associated with a data set such that the first feature is associated with a first data set and the second feature is associated with a second data set,

the first data set associated with the first feature comprises values for the first feature,

the second data set associated with the second feature comprises values for the second feature, and

the method comprises:

for each feature included in the set of features, obtaining a cost value for the feature, the cost value indicating a cost of providing the data set associated with the feature from a source node that stores the data set to the destination node that is connected to the source node via a network;

for each feature included in the set of features, obtaining an importance value indicating an importance of the feature;

for each feature included in the set of features, calculating a priority value using the importance value for the feature and the cost value for the feature;

based on the calculated priority values, selecting a subset of the set of features; and

for each selected feature, initiating the transmission of the respective data set associated with the respective selected feature from the source node that stores the respective data set associated with the respective selected feature to the destination node via the network that connects the destination node with the source node that stores the respective data set associated with the respective selected feature, wherein

the destination node is configured to execute a machine learning process that is operable to use the respective data sets to produce a model, and/or

the destination node is configured to execute a model generated by a machine learning process, and the model is operable to use the respective data sets to produce an inference.

2 . The method of claim 1 , wherein

the step of obtaining the cost value for each feature included in the set of features comprises obtaining a first cost value for the first feature, and

the first cost value is a function of size of the first data set.

3 . The method of claim 2 , wherein

the method further comprises obtaining network performance information for the network that connects the destination node with the source node that stores the first data set associated with the first feature, and

the first cost value is also the function of network performance information.

4 . The method of claim 1 , wherein the first cost value is inversely proportional to the feature importance value.

5 . The method of claim 1 , wherein

the subset of features includes the first feature but not the second feature, and

the method further comprises:

after initiating the transmission of the data sets for the selected features, receiving from the destination node a model training result pertaining to a training of a model using the first data set associated with the first feature; and

after receiving the model training result, initiating the transmission of the second data set associated with the second feature from the source node that stores the second data set associated with the second feature to the destination node.

6 . The method of claim 5 , wherein the model training result comprises a first feature importance value for the first feature, the first feature importance value indicating the first feature's impact on the performance of the model.

7 . A non-transitory computer readable storage medium storing a computer program comprising instructions which when executed by processing circuitry causes the processing circuitry to perform the method of claim 1 .

8 . An apparatus for providing data to a destination node connected to one or more source nodes via one or more networks and for selecting at least one feature from a set of features, the set of features comprising a first feature and a second feature, wherein

each feature included in the set of features is associated with a data set such that the first feature is associated with a first data and the second feature is associated with a second data set,

the first data set comprises values for the first feature,

the second data comprises values for the second feature, and

the apparatus including:

a memory; and

processing circuitry, wherein the apparatus is configured to:

for each feature included in the set of features, obtain a cost value for the feature, the cost value indicating a cost of providing the data set associated with the feature from a source node that stores the data set to the destination node that is connected to the source node via a network;

for each feature included in the set of features, obtain an importance value indicating an importance of the feature;

for each feature included in the set of features, calculate a priority value using the importance value for the feature and the cost value for the feature;

based on the calculated priority values, select a subset of the set of features; and

for each selected feature, initiate the transmission of the respective data set associated with the respective selected feature from the source node that stores the respective data set associated with the respective selected feature to the destination node via the network that connects the destination node with the source node that stores the respective data set associated with the respective selected feature, wherein the destination node is configured to:

execute a machine learning process that is operable to use the respective data sets to produce a model, and/or

execute a model generated by a machine learning process, and the model is operable to use the respective data sets to produce an inference.

9 . The apparatus of claim 8 , wherein

the apparatus is further configured to obtain the value for each feature included in the set of features by performing a process that includes obtaining a first cost value for the first feature, and

the first cost value is a function of size of the first data set.

10 . The apparatus of claim 8 , wherein

the apparatus is further configured to obtain network performance information for the network that connects the destination node with the source node that stores the first data set associated with the first feature, and

the first cost value is also a function of network performance information.

11 . The apparatus of claim 8 , wherein the first cost value is inversely proportional to the feature importance value.

12 . The apparatus of claim 8 , wherein

the subset of features includes the first feature but not the second feature, and

the apparatus is further configured to:

after initiating the transmission of the data sets for the selected features, receive from the destination node a model training result pertaining to a training of a model using the first data set associated with the first feature; and

after receiving the model training result, initiate the transmission of the second data set associated with the second feature from the source node that stores the second data set associated with the second feature to the destination node.

13 . The apparatus of claim 12 , wherein the model training result comprises a first feature importance value for the first feature, the first feature importance value indicating the first feature's impact on the performance of the model.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 23, 2022
From: FERSMAN, ELENA; FOLKESSON, MATS; INAM, RAFIA; LINDSTRÖM, DANIEL; SUN, BIN
To: TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)
Reel/Frame 061197/0701 →
Continuity (1)
Related Publication 20230075551A1 · Mar 9, 2023
References Cited (23)
US 6963870B2 · Heckerman · 2005 [cited by applicant]
US 11640447B1 · Stone · 2023 [cited by examiner]
US 20060274663A1 · Brethereau · 2006 [cited by examiner]
US 20160094649A1 · Dornquast · 2016 [cited by examiner]
US 20170149863A1 · Chitti · 2017 [cited by examiner]
US 20180365521A1 · Dai · 2018 [cited by examiner]
US 20190156215A1 · Matveev et al. · 2019 [cited by applicant]
US 20190220703A1 · Prakash et al. · 2019 [cited by applicant]
US 20190318270A1 · Duenner et al. · 2019 [cited by applicant]
US 20210073672A1 · Shi · 2021 [cited by examiner]
CN 108768876A · 2018 [cited by applicant]
CN 109558909A · 2019 [cited by applicant]
CN 110191053A · 2019 [cited by applicant]
CN 110659740A · 2020 [cited by applicant]
EP 3550469A2 · 2019 [cited by applicant]
Dongzhu Liu, et al. (Liu) Data-Importance Aware User Scheduling for Communication Efficient Edge Machine Learning, Oct. 5, 2019, Https://arxiv.org/abs/1910.02214 (it's the same as the provided NPL) (Year: 2019). [cited by examiner]
Inagaki et al. “Prioritization of Mobile IoT Data Transmission Based on Data Importance Extracted From Machine Learning Model” (IEEE, IEEE Access, vol. 7, date of publication Jul. 11, 2019, DOI: 10.1109/Access.2019.2928… [cited by examiner]
International Search Report and Written Opinion issued in International Application No. PCT/SE2020/050122 dated Feb. 4, 2021 (14 pages). [cited by applicant]
Liu, D., et al., “Data-Importance Aware User Scheduling for Communication-Efficient Edge Machine Learning”, arXiv:1910.02214v1 [cs.NI], Oct. 5, 2019 (31 pages). [cited by applicant]
Lyu, X., et al., “Optimal Online Data Partitioning for Geo-Distributed Machine Learning in Edge of Wireless Network”, IEEE Journal on Selected Areas in Communications, vol. 37, No. 10, Oct. 2019 (14 pages). [cited by applicant]
Yu, W. et al., “Research of Datasets Division in Abstract Grid Workflow”, Aug. 10, 2013 (4 pages). [cited by applicant]
Qiao Lan et al., “An Introduction to Communication Efficient Edge Machine Learning”, Dec. 2019 (22 pages). [cited by applicant]
Mona Shah et al., “Prognosis Using Distributed Data Classification with Privacy Preserving: A Novel Approach”, IEEE, Dec. 19-21, 2016 (4 pages). [cited by applicant]