IP Library Granted Patent US 12,353,455
Granted Patent B2
US 12,353,455 · App. 18/621,820 · Granted Jul 8, 2025

Search results for pseudo-content

Inventors: Matthew Allen Strong Ross (London, CA); Azadeh Haji Hosseini (Toronto, CA); Monique Alves Cruz (Toronto, CA); Albert Jimenez Sanfiz (Toronto, CA); Prabhdeep Singh Cheema (Dublin, CA); Hima Kiran Alladi (Missouri City, TX)
Assignee: Scribd, Inc.
G06F16/334G06F16/335
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,353,455
App. No.
18/621,820
Granted
Jul 8, 2025
Kind
B2
Abstract

Various embodiments of a Title Engine generate search results that identify content available in a content corpus in response to receiving a search query for content that is currently unavailable in the content corpus. Rather than returning output merely indicating absence of the requested content set forth in the received search query, the Title Engine identifies various content available in the content corpus that is similar to the search query's requested—but unavailable—content. The Title Engine identifies content in a content corpus that is similar to requested content that has been determined as unavailable. Upon determining unavailability of the requested particular content in the content corpus, the Title Engine generates a pseudo-identifier for the requested particular content. The Title Engine inserts the pseudo-identifier into a sequence of content identifiers. The Title Engine generates an embedding for the pseudo-identifier.

Claims (59)

1. A computer-implemented method, comprising:

receiving a search query, from a user account, for content submitted to a content corpus;

determining the content identified in the search query is unavailable;

generating a pseudo-identifier for the unavailable content identified in the search query;

inserting the pseudo-identifier into a sequence of content identifiers, the sequence associated with the user account;

generating an embedding for the pseudo-identifier;

identifying, via the embedding, a plurality of sequences of content identifiers, wherein a respective sequence corresponds to a different user account that previously sent a search query for the unavailable content identified by the pseudo-identifier;

for one or more of the respective sequences identified via the embedding:

(i) determining content interactions that occurred within a particular time range from an instance of the pseudo-identifier placed in the respective sequence, each of the determined content interactions associated with a proximate content identifier; and

(ii) identifying neighbor content from at least a portion of the proximate content identifiers;

generating search results responsive to the user account's search query based on the neighbor content.

2. The computer-implemented method of claim 1 , wherein each content identifier uniquely identifies a different portion of content accessed by the user account; and

wherein the sequence of content identifiers orders the content identifiers according to an access time by the user account of each respective content identifier.

3. The computer-implemented method of claim 1 , further comprising:

generating, via the embedding, search results responsive to the search query from the user account, the search results including at least one recommendation of similar content, the similar content comprising content available in the content corpus and similar to the unavailable content identified in the search query.

4. The computer-implemented method of claim 1 , wherein identifying a plurality of sequences of content identifiers further comprises:

identifying previous matching search queries from one or more of the different user accounts that match the search query from the user account within a threshold of similarity;

identifying respective different pseudo-identifiers that correspond to the previous matching search queries; and

identifying a plurality of sequence content identifiers, wherein a respective sequence includes an instance of at least one of the respective different pseudo-identifiers.

5. The computer-implemented method of claim 1 , wherein at least a portion of the content corpus comprises content uploaded from a first user account for access by a plurality of user accounts.

6. A system comprising one or more processors, and a non-transitory computer-readable medium including one or more sequences of instructions that, when executed by the one or more processors, cause the system to perform operations comprising:

receiving a search query, from a user account, for content submitted to a content corpus;

determining the content identified in the search query is unavailable;

generating a pseudo-identifier for the unavailable content identified in the search query;

inserting the pseudo-identifier into a sequence of content identifiers, the sequence associated with the user account;

generating an embedding for the pseudo-identifier;

identifying, via the embedding, a plurality of sequences of content identifiers, wherein a respective sequence corresponds to a different user account that previously sent a search query for the unavailable content identified by the pseudo-identifier;

for one or more of the respective sequences identified via the embedding:

(i) determining content interactions that occurred within a particular time range from an instance of the pseudo-identifier placed in the respective sequence, each of the determined content interactions associated with a proximate content identifier; and

(ii) identifying neighbor content from at least a portion of the proximate content identifiers;

generating search results responsive to the user account's search query based on the neighbor content.

7. The system of claim 6 , wherein each content identifier uniquely identifies a different portion of content accessed by the user account; and

wherein the sequence of content identifiers orders the content identifiers according to an access time by the user account of each respective content identifier.

8. The system of claim 6 , further comprising:

generating, via the embedding, search results responsive to the search query from the user account, the search results including at least one recommendation of similar content, the similar content comprising content available in the content corpus and similar to the unavailable content identified in the search query.

9. The system of claim 6 , wherein identifying a plurality of sequences of content identifiers further comprises:

identifying previous matching search queries from one or more of the different user accounts that match the search query from the user account within a threshold of similarity;

identifying respective different pseudo-identifiers that correspond to the previous matching search queries; and

identifying a plurality of sequence content identifiers, wherein a respective sequence includes an instance of at least one of the respective different pseudo-identifiers.

10. The system of claim 6 , wherein at least a portion of the content corpus comprises content uploaded from a first user account for access by a plurality of user accounts.

11. A computer program product comprising a non-transitory computer-readable medium having a computer-readable program code embodied therein to be executed by one or more processors, the program code including instructions to:

receiving a search query, from a user account, for content submitted to a content corpus;

determining the content identified in the search query is unavailable;

generating a pseudo-identifier for the unavailable content identified in the search query;

inserting the pseudo-identifier into a sequence of content identifiers, the sequence associated with the user account;

generating an embedding for the pseudo-identifier;

identifying, via the embedding, a plurality of sequences of content identifiers, wherein a respective sequence corresponds to a different user account that previously sent a search query for the unavailable content identified by the pseudo-identifier;

for one or more of the respective sequences identified via the embedding:

(i) determining content interactions that occurred within a particular time range from an instance of the pseudo-identifier placed in the respective sequence, each of the determined content interactions associated with a proximate content identifier; and

(ii) identifying neighbor content from at least a portion of the proximate content identifiers;

generating search results responsive to the user account's search query based on the neighbor content.

12. The computer program product of claim 11 , wherein each content identifier uniquely identifies a different portion of content accessed by the user account; and

wherein the sequence of content identifiers orders the content identifiers according to an access time by the user account of each respective content identifier.

13. The computer program product of claim 11 , further comprising:

generating, via the embedding, search results responsive to the search query from the user account, the search results including at least one recommendation of similar content, the similar content comprising content available in the content corpus and similar to the unavailable content identified in the search query.

14. The computer program product of claim 11 , wherein identifying a plurality of sequences of content identifiers further comprises:

identifying previous matching search queries from one or more of the different user accounts that match the search query from the user account within a threshold of similarity;

identifying respective different pseudo-identifiers that correspond to the previous matching search queries; and

identifying a plurality of sequence content identifiers, wherein a respective sequence includes an instance of at least one of the respective different pseudo-identifiers.

Assignments (3)
RELEASE OF SECURITY INTEREST Recorded Jul 9, 2026
From: CITIBANK, N.A.
To: SCRIBD, INC.
Reel/Frame 075216/0990 →
SECURITY INTEREST Recorded Apr 2, 2024
From: SCRIBD, INC.
To: CITIBANK, N.A.
Reel/Frame 066984/0345 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 29, 2024
From: ROSS, MATTHEW ALLEN STRONG; HAJI HOSSEINI, AZADEH; ALVES CRUZ, MONIQUE; JIMENEZ SANFIZ, ALBERT; CHEEMA, PRABHDEEP SINGH; ALLADI, HIMA KIRAN
To: SCRIBD, INC.
Reel/Frame 066951/0001 →
Continuity (3)
Continuation In Part 18117271 · Mar 3, 2023
Provisional Application 63455775 · Mar 30, 2023
Related Publication 20240296179A1 · Sep 5, 2024
References Cited (14)
US 8438469B1 · Scott et al. · 2013 [cited by applicant]
US 8515908B2 · Kumar · 2013 [cited by examiner]
US 8688669B1 · Bernstein · 2014 [cited by examiner]
US 9444940B2 · Skiba · 2016 [cited by examiner]
US 10146852B1 · Chu · 2018 [cited by examiner]
US 11294974B1 · Shukla et al. · 2022 [cited by applicant]
US 20070016848A1 · Rosenoff et al. · 2007 [cited by applicant]
US 20080222125A1 · Chowdhury · 2008 [cited by examiner]
US 20130054583A1 · Macklem · 2013 [cited by examiner]
US 20140316890A1 · Kagan · 2014 [cited by examiner]
US 20150161192A1 · Scoles · 2015 [cited by examiner]
US 20150161202A1 · Shapira · 2015 [cited by examiner]
US 20200004886A1 · Ramanath · 2020 [cited by examiner]
Extended European Search Report received in European application No. 24167943.0, mailed on Aug. 6, 2024. [cited by applicant]
Cited By (1)
US 12,572,565