IP Library Granted Patent US 12,456,143
Granted Patent B2
US 12,456,143 · App. 17/323,652 · Granted Oct 28, 2025

Enhanced neural network systems and methods

Inventors: Geoffrey B. Rhoads (West Linn, OR); Ravi K. Sharma (Portland, OR)
Assignee: Digimarc Corporation
G06Q30/0641G06F18/214G06F18/22G06F18/2431G06V10/764G06V10/774G06N3/04G06N3/08
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,456,143
App. No.
17/323,652
Granted
Oct 28, 2025
Kind
B2
Abstract

Two stages of a convolutional neural network are linked by an interconnect that effects a spatial transposition of array data. The spatial transposition can include rotation, scaling, or translation (e.g., in x- or y-directions). A parameter characterizing the transposition (e.g., a parameter identifying rotation angle) can be learned by the same training process that is also used to learn other network parameters, such as layer coefficients. Additionally, or alternatively, data input to a neural network comprises—for each pixel in a patch of imagery—plural data that each indicates a relationship between the value of the pixel, and the value of a neighboring pixel. Some such neural networks can be trained to indicate the presence of a digital watermark signal in the patch of imagery—or a parameter characterizing such a digital watermark signal. Other features and arrangements are also detailed.

Claims (40)

1. In a neural network comprising first and second stages configured for image recognition, at least one of said stages including a convolutional layer, and a method comprising the acts:

receiving an array of output data from the first stage, said array of output data having dimensions of at least 11×11 elements; and

spatially-transposing said array of output data, and employing a resulting array of spatially-transposed data as input data in the second stage;

wherein spatially-transposing the array of output data prior to provision to the second stage, and

wherein the output of the first stage is coupled to the input of the second stage by a twisted coupling interconnect to increase the reliability of image recognition at different orientations.

2. The method of claim 1 in which said spatially-transposing comprises rotating.

3. The method of claim 1 in which said spatially-transposing comprises scaling.

4. The method of claim 1 in which said spatially-transposing comprises spatial translating.

5. The method of claim 1 in which said spatially-transposing comprises rotating, scaling, and spatially-translating.

6. The method of claim 1 that includes learning a parameter that characterizes said operation through a training process.

7. The method of claim 6 in which said training process is a gradient descent training process.

8. The method of claim 6 that further includes learning coefficients for one or more of said stages by said training process.

9. The method of claim 1 in which said spatially-transposing includes computing one value in said array of spatially-transposed data as a weighted sum of plural values in said array of output data.

10. The method of claim 9 in which said array of output data includes plural rows of data including a top row, a bottom row, and plural intermediate rows, wherein said weighted sum of plural values is computed in toroidal fashion, as a sum that includes one or more values from said top row, and one or more values from said bottom row, but no value from one of said intermediate rows.

11. The method of claim 1 in which spatially-transposing the array of output data comprises one or more operations from the list: rotating, scaling and spatially-translating.

12. The method of claim 11 in which spatially-transposing the array of output data comprises two of: rotating, scaling, and spatially-translating.

13. A neural network system for image recognition, comprising:

a first stage and a second stage, at least one of said stages including a convolutional layer; and

a multi-core processor configured to:

receive an array of output data from the first stage, said array of output data having dimensions of at least 11×11 elements;

spatially-transpose said array of output data; and

employ a resulting array of spatially-transposed data as input data in the second stage,

wherein the output of the first stage is coupled to the input of the second stage by a twisted coupling interconnect to increase the reliability of image recognition at different orientations.

14. The system of claim 13 , wherein said multi-core processor is further configured to learn a parameter that characterizes said operation through a training process.

15. The system of claim 14 , wherein said training process is a gradient descent training process.

16. The system of claim 14 , wherein said multi-core processor is further configured to learn coefficients for one or more of said stages by said training process.

17. The system of claim 13 , wherein said spatially-transposing includes computing one value in said array of spatially-transposed data as a weighted sum of plural values in said array of output data.

18. The system of claim 17 , wherein said array of output data includes plural rows of data including a top row, a bottom row, and plural intermediate rows, wherein said weighted sum of plural values is computed in toroidal fashion, as a sum that includes one or more values from said top row, and one or more values from said bottom row, but no value from one of said intermediate rows.

19. The system of claim 13 , wherein the second stage comprises a digital watermarking stage, and said spatially-transposed data is utilized in the digital watermark stage to resolve image distortion.

20. A neural network system for image recognition, comprising:

a first stage, having an input and an output, for processing first input data applied to the input to produce first output data at the output, said first output data comprising array data having dimensions of at least 11×11 elements; and

a second stage having an input and an output, for processing second input data applied to the second stage input to produce second output data at the second stage output, at least one of said first stage and said second stage including a convolutional layer;

wherein the output of the first stage is coupled to the input of the second stage by a twisted coupling interconnect;

wherein said array data produced at the output of the first stage is spatially-transposed prior to being applied to the input of the second stage as the second input data to increase the reliability of image recognition at different orientations.

21. The system of claim 20 , further comprising a processor configured to learn a parameter that characterizes said operation through a training process.

22. The system of claim 21 , wherein said training process is a gradient descent training process.

23. The system of claim 21 , further comprising a processor configured to learn coefficients for one or more of said stage means by said training process.

24. The system of claim 20 , wherein said spatially transposed data is produced by computing one value as a weighted sum of plural values in said array of output data.

25. The system of claim 24 , wherein said array of output data includes plural rows of data including a top row, a bottom row, and plural intermediate rows, further comprising a processor that computes said weighted sum of plural values in toroidal fashion, as a sum that includes one or more values from said top row, and one or more values from said bottom row, but no value from one of said intermediate rows.

26. The system of claim 20 , wherein the second stage comprises digital watermarking, and said spatially-transposed data is utilized by said digital watermarking to resolve image distortion.

Assignments (2)
ARTICLES OF CONVERSION Recorded Jun 19, 2026
From: DIGIMARC CORPORATION
To: DIGIMARC LLC
Reel/Frame 075863/0211 →
ARTICLES OF AMENDMENT OFTHE ARTICLES OF ORGANIZATION OF DIGIMARC LLC Recorded Jun 19, 2026
From: DIGIMARC LLC
To: DMRC LLC
Reel/Frame 075863/0266 →
Continuity (5)
Continuation In Part 16880778 · May 21, 2020
Division 15726290 · Oct 5, 2017
Provisional Application 63029662 · May 25, 2020
Provisional Application 62404721 · Oct 5, 2016
Related Publication 20210357690A1 · Nov 18, 2021
References Cited (46)
US 5903884A · Lyon · 1999 [cited by applicant]
US 10007863B1 · Pereira · 2018 [cited by applicant]
US 10664722B1 · Sharma · 2020 [cited by applicant]
US 20130308045A1 · Rhoads · 2013 [cited by applicant]
US 20140108309A1 · Frank · 2014 [cited by applicant]
US 20150055855A1 · Rodriguez · 2015 [cited by applicant]
US 20150161482A1 · Preetham · 2015 [cited by applicant]
US 20150170001A1 · Rabinovich · 2015 [cited by applicant]
US 20150242707A1 · Wilf · 2015 [cited by applicant]
US 20150278224A1 · Jaber · 2015 [cited by applicant]
US 20160140425A1 · Kulkarni · 2016 [cited by applicant]
US 20160267358A1 · Shoaib · 2016 [cited by applicant]
US 20160379091A1 · Lin · 2016 [cited by applicant]
US 20170132496A1 · Shoaib · 2017 [cited by examiner]
US 20170294010A1 · Shen · 2017 [cited by applicant]
US 20170316281A1 · Criminisi · 2017 [cited by applicant]
US 20180005343A1 · Rhoads · 2018 [cited by applicant]
US 20180032844A1 · Yao · 2018 [cited by applicant]
US 20180165548A1 · Wang · 2018 [cited by applicant]
US 20190057314A1 · Julian · 2019 [cited by applicant]
US 20190266749A1 · Rhoads · 2019 [cited by applicant]
CN 105740909A · 2016 [cited by examiner]
NPL—Bai et al., (CN 105740909 A) Published Jul. 6, 2016 (Machine Translation total 12 pages) (Year: 2016). [cited by examiner]
Coli, et al, The toroidal neural networks, 2000 IEEE International Symposium on Circuits and Systems (ISCAS) May 28, 2000 (vol. 4, pp. 137-140). [cited by applicant]
Palazzari, et al, Massively parallel processing implementation of the toroidal neural networks, Proceedings of the 2000 6th IEEE International Workshop on Cellular Neural Networks and their Applications, May 25, 2000 (p… [cited by applicant]
Mikulski, et al, Toroidal AutoEncoder, arXiv preprint arXiv:1903.12286. Mar. 28, 2019. [cited by applicant]
Chapelle, et al, Model selection for support vector machines, Advances in neural information processing systems. 1999;12:230-6. [cited by applicant]
Tang, et al, Diagonal and toroidal mesh networks, IEEE Transactions on Computers. Jul. 1994;43(7):815-26. [cited by applicant]
Babenko, et al, Neural Codes for Image Retrieval, arXiv preprint, arXiv:1404.1777 (2014), 16 pages. [cited by applicant]
Krizhevsky, et al, Imagenet Classification with Deep Convolutional Neural Networks, Advances in Neural Information Processing Systems, 2012, 9 pages. [cited by applicant]
LeCun et al, Handwritten Digit Recognition with a Back-Propagation Network, Advances in Neural Information Processing Systems, pp. 396-404, 1990. [cited by applicant]
Szegedy, et al, Going Deeper with Convolutions. IEEE Conference on Computer Vision and Pattern Recognition, pp. 1-9, 2015. [cited by applicant]
Szegedy, et al, Intriguing Properties of Neural Networks, arXiv preprint, arXiv:1312.6199v4, 2014, 10 pages. [cited by applicant]
Zeiler et al, Visualizing and Understanding Convolutional Networks, European Conference on Computer Vision, 2014, 16 pages. [cited by applicant]
Chen et al, Clusternet: Deep hierarchical cluster network with rigorously rotation-invariant representation for point cloud analysis, IEEE/CVF Conference on Computer Vision and Pattern Recognition 2019 (pp. 4994-5002). [cited by applicant]
Cheng et al, Rifd-CNN: Rotation-invariant and fisher discriminative convolutional neural networks for object detection, IEEE Conference on Computer Vision and Pattern Recognition 2016 (pp. 2884-2893). [cited by applicant]
Chidester, et al, Rotation equivariance and invariance in convolutional neural networks, arXiv preprint arXiv:1805.12301, May 31, 2018. [cited by applicant]
Deng, et al, Joint hand detection and rotation estimation using CNN, IEEE Transactions on Image Processing, Dec. 4, 2017;27(4):1888-900. [cited by applicant]
Esteves, et al, Polar transformer networks, arXiv preprint arXiv:1709.01889, Sep. 6, 2017. [cited by applicant]
Follmann, et al, A rotationally-invariant convolution module by feature map back-rotation, 2018 IEEE Winter Conference on Applications of Computer Vision (WACV) Mar. 12, 2018 (pp. 784-792), IEEE. [cited by applicant]
Jaderberg et al, Spatial transformer networks, arXiv preprint arXiv:1506.02025, Jun. 5, 2015. [cited by applicant]
Ma et al, Arbitrary-oriented scene text detection via rotation proposals, IEEE Transactions on Multimedia, Mar. 23, 2018;20(11):3111-22. [cited by applicant]
Marcos et al, Rotation equivariant vector field networks, IEEE International Conference on Computer Vision 2017 (pp. 5048-5057). [cited by applicant]
Shi et al, Real-time rotation-invariant face detection with progressive calibration networks, IEEE Conference on Computer Vision and Pattern Recognition 2018 (pp. 2295-2303). [cited by applicant]
Weiler et al, Learning steerable filters for rotation equivariant CNNs, IEEE Conference on Computer Vision and Pattern Recognition 2018 (pp. 849-858). [cited by applicant]
Zhang X, Liu L, Xie Y, Chen J, Wu L, Pietikainen M, Rotation invariant local binary convolution neural networks, IEEE International Conference on Computer Vision Workshops 2017 (pp. 1210-1219). [cited by applicant]