IP Library Granted Patent US 11,651,160
Granted Patent B2
US 11,651,160 · App. 16/739,655 · Granted May 16, 2023

Systems and methods for using machine learning and rules-based algorithms to create a patent specification based on human-provided patent claims such that the patent specification is created without human intervention

Inventors: Ian C. Schick (Ojai, CA); Kevin Knight (Marina del Rey, CA)
Assignee: Specifio, Inc.
G06F40/30G06F40/137G06F40/284G06F40/56G06Q10/10G06Q50/184G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,651,160
App. No.
16/739,655
Granted
May 16, 2023
Kind
B2
Abstract

Systems and methods for using machine learning and rules-based algorithms to create a patent specification based on human-provided patent claims such that the patent specification is created without human intervention are disclosed. Exemplary implementations may: obtain a claim set; obtain a first data structure representing the claim set; obtain a second data structure; obtain a third data structure; and determine one or more sections of the patent specification based on the first data structure, the second data structure, and the third data structure.

Claims (33)

1. A system configured to automatically convert patent claim language into prose, the system comprising:

one or more hardware processors configured by machine-readable instructions to:

obtain at least a portion of a patent claim, the patent claim being a numbered sentence that precisely defines an invention;

determine, based on the at least the portion of the patent claim, a first data structure representing the at least the portion of the patent claim, the first data structure including a first data structure element, the first data structure element including one or more language units from the at least the portion of the patent claim, wherein an individual language unit of the first data structure includes one or more clauses;

tokenize the first data structure element into tokens by breaking up a stream of text in the first data structure to generate a tokenized first data structure element, wherein a given token includes a word, a phrase, a symbol, or a punctuation;

obtain, through a first machine learning model, a parse of the tokenized first data structure element, the parse of the tokenized first data structure element including one or more of grammatical constituents, parts of speech, syntactic relations, or inflectional form; and

determine, based on one or both of the at least the portion of the patent claim or the first data structure, a second data structure, the second data structure including a second data structure element, the second data structure element including one or more language units associated with the at least the portion of the patent claim, an individual language unit of the second data structure being in prose, wherein prose includes an ordinary form of written language, wherein the individual language unit in the first data structure is transformed into a corresponding language unit in the second data structure based on a computerized natural language generation operation using one or both of a second machine learning model or rules-based algorithms.

2. The system of claim 1 , wherein the patent claim is part of a claim set, the claim set including an independent claim and zero or more dependent claims, each dependent claim in the claim set depending on the independent claim by referring to the independent claim or an intervening dependent claim.

3. The system of claim 1 , wherein the one or more language units in the first data structure are organized according to one or more classifications of individual language units, wherein the one or more classifications include one or more of independent claim, dependent claim, preamble, main feature, sub feature, claim line, clause, phrase, or word.

4. The system of claim 1 , wherein the first data structure and the second data structure are separate and distinct from each other.

5. The system of claim 1 , wherein the computerized natural language generation operation includes one or more of paraphrase induction, simplification, compression, clause fusion, or expansion.

6. The system of claim 1 , wherein the patent claim is part of a claim set, wherein the one or more hardware processors are further configured by the machine-readable instructions to determine, based on the one or more language units from the first data structure and/or the one or more language units from the second data structure, a third data structure, wherein ordered content of the third data structure is ordered based on one or more of claim structure of the claim set, antecedent basis in the claim set, or claim dependency in the claim set.

7. The system of claim 6 , wherein one or more sections of a patent specification are further determined by assembling language units from the third data structure.

8. The system of claim 1 , wherein an individual data structure includes a specialized format for organizing and storing data, the individual data structure including one or more of an array, a list, two or more linked lists, a stack, a queue, a graph, a table, or a tree.

9. The system of claim 1 , wherein the one or more hardware processors are further configured by the machine-readable instructions to generate one or more drawing figures described in a patent specification, the one or more drawing figures being generated based on one or both of the first data structure or the second data structure.

10. A computer-implemented method to automatically convert patent claim language into prose, the method comprising:

obtaining at least a portion of a patent claim, the patent claim being a numbered sentence that precisely defines an invention;

determining, based on the at least the portion of the patent claim, a first data structure representing the at least the portion of the patent claim, the first data structure including a first data structure element, the first data structure element including one or more language units from the at least the portion of the patent claim, wherein an individual language unit of the first data structure includes one or more clauses;

tokenizing the first data structure element into tokens by breaking up a stream of text in the first data structure to generate a tokenized first data structure element, wherein a given token includes a word, a phrase, a symbol, or a punctuation;

obtaining, through a first machine learning model, a parse of the tokenized first data structure element, the parse of the tokenized first data structure element including one or more of grammatical constituents, parts of speech, syntactic relations, or inflectional form; and

determining, based on one or both of the at least the portion of the patent claim or the first data structure, a second data structure, the second data structure including a second data structure element, the second data structure element including one or more language units associated with the at least the portion of the patent claim, an individual language unit of the second data structure being in prose, wherein prose includes an ordinary form of written language, wherein the individual language unit in the first data structure is transformed into a corresponding language unit in the second data structure based on a computerized natural language generation operation using one or both of a second machine learning model or rules-based algorithms.

11. The method of claim 10 , wherein the patent claim is part of a claim set, the claim set including an independent claim and zero or more dependent claims, each dependent claim in the claim set depending on the independent claim by referring to the independent claim or an intervening dependent claim.

12. The method of claim 10 , wherein the one or more language units in the first data structure are organized according to one or more classifications of individual language units, wherein the one or more classifications include one or more of independent claim, dependent claim, preamble, main feature, sub feature, claim line, clause, phrase, or word.

13. The method of claim 10 , wherein the first data structure and the second data structure are separate and distinct from each other.

14. The method of claim 10 , wherein the computerized natural language generation operation includes one or more of paraphrase induction, simplification, compression, clause fusion, or expansion.

15. The method of claim 10 , wherein the patent claim is part of a claim set, and wherein the method further comprises determining, based on the one or more language units from the first data structure and/or the one or more language units from the second data structure, a third data structure, wherein ordered content of the third data structure is ordered based on one or more of claim structure of the claim set, antecedent basis in the claim set, or claim dependency in the claim set.

16. The method of claim 15 , wherein one or more sections of a patent specification are further determined by assembling language units from the third data structure.

17. The method of claim 10 , wherein an individual data structure includes a specialized format for organizing and storing data, the individual data structure including one or more of an array, a list, two or more linked lists, a stack, a queue, a graph, a table, or a tree.

18. The method of claim 10 , further comprising generating one or more drawing figures described in a patent specification, the one or more drawing figures being generated based on one or both of the first data structure or the second data structure.

19. The system of claim 1 , wherein the one or more hardware processors are further configured by the machine-readable instructions to:

determine at least a portion of one or more sections of a patent specification by assembling the one or more language units from the first data structure and the one or more language units from the second data structure.

20. The method of claim 10 , further comprising:

determining at least a portion of one or more sections of a patent specification by assembling the one or more language units from the first data structure and the one or more language units from the second data structure.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2024
From: SPECIFIO, INC.
To: PAXIMAL, INC.
Reel/Frame 066266/0834 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 10, 2020
From: SCHICK, IAN C.; KNIGHT, KEVIN
To: SPECIFIO, INC.
Reel/Frame 051478/0974 →
Continuity (8)
Continuation 16510074 · Jul 12, 2019
Continuation 15892679 · Feb 9, 2018
Provisional Application 62459199 · Feb 15, 2017
Provisional Application 62459235 · Feb 15, 2017
Provisional Application 62459357 · Feb 15, 2017
Provisional Application 62459208 · Feb 15, 2017
Provisional Application 62459246 · Feb 15, 2017
Related Publication 20200151393A1 · May 14, 2020
Cited By (2)
US 12,662,012 US 12,688,367