IP Library › Granted Patent US 12,210,567
Granted Patent B2
US 12,210,567 · App. 18/071,986 · Granted Jan 28, 2025

Methods, systems, and media for determining playlist title coherence and quality

Inventors: Ben Goodrich (San Francisco, CA); Kumar Chippala (Mountain View, CA)
Assignee: Google LLC
G06F16/639G06F40/263G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,210,567
App. No.
18/071,986
Granted
Jan 28, 2025
Kind
B2
Abstract

Methods, systems, and media for determining playlist title coherence and quality are provided. In some embodiments, a method for generating playlist recommendations includes: determining, using a hardware processor, a title of a playlist; generating, using the hardware processor, a byte-level representation of the title based on the title of the playlist; determining, using the hardware processor, an embedded representation of the title based on the byte-level representation; determining, using the hardware processor, a perplexity score of the title by inputting the embedded representation of the title into a trained language model, wherein the perplexity score is an output of the trained language model; and causing, using the hardware processor, a recommendation based on the perplexity score of the title to be presented.

Claims (40)

1. A method comprising:

detecting, using a hardware processor, a user-generated title associated with a playlist available on a streaming service;

generating, using the hardware processor, a byte-level representation of the user-generated title based on a sequence of bytes corresponding to the user-generated title of the playlist;

determining, using the hardware processor, an embedded representation of the user-generated title based on the byte-level representation;

determining, using the hardware processor, a plurality of perplexity scores that each correspond to an individual perplexity score of a portion of the embedded representation of the user-generated title by inputting the embedded representation of the user-generated title into a trained language model, wherein the plurality of perplexity scores is an output of the trained language model;

determining, using the hardware processor, an overall perplexity score of the user-generated title based on the plurality of perplexity scores, a number of bytes in the byte-level representation, and a position of each byte in the byte-level representation, wherein the overall perplexity score indicates whether the sequence of bytes corresponding to the user-generated title of the playlist is an unlikely sequence of bytes or a likely sequence of bytes; and

responsive to determining that the overall perplexity score indicates that the sequence of bytes corresponding to the user-generated title of the playlist is a likely sequence of bytes, causing, using the hardware processor, a recommendation system associated with the streaming service to recommend the playlist based on the overall perplexity score of the user-generated title.

2. The method of claim 1 , wherein the trained language model is trained using a plurality of training data that includes at least one of a plurality of pre-approved professionally curated playlist titles and a plurality of pre-approved professionally curated playlist descriptions.

3. The method of claim 1 , wherein the trained language model is trained using a plurality of training data that includes a plurality of pre-approved professionally curated playlist titles, and wherein the trained language model is configured to generate the plurality of perplexity scores of the user-generated title of the playlist based on a negative log-likelihood that the sequence of bytes corresponding to the user-generated title of the playlist would appear in one of the plurality of pre-approved professionally curated playlist titles in the plurality of training data.

4. The method of claim 1 , further comprising:

responsive to determining that the overall perplexity score indicates that the sequence of bytes corresponding to the user-generated title is an unlikely sequence of bytes, causing, using the hardware processor, the recommendation system associated with the streaming service to exclude the playlist from being recommended.

5. The method of claim 1 , further comprising determining that the user-generated title is based in a particular language from a plurality of languages.

6. The method of claim 1 , wherein the playlist comprises a plurality of music tracks.

7. A system for generating playlist recommendations, the system comprising:

a hardware processor that is configured to:

detect a user-generated title associated with a playlist available on a streaming service;

generate a byte-level representation of the user-generated title based on a sequence of bytes corresponding to the user-generated title of the playlist;

determine an embedded representation of the user-generated title based on the byte-level representation;

determine a plurality of perplexity scores that each correspond to an individual perplexity score of a portion of the embedded representation of the user-generated title by inputting the embedded representation of the user-generated title into a trained language model, wherein the plurality of perplexity scores is an output of the trained language model;

determine an overall perplexity score of the user-generated title based on the plurality of perplexity scores, a number of bytes in the byte-level representation, and a position of each byte in the byte-level representation, wherein the overall perplexity score indicates whether the sequence of bytes corresponding to the user-generated title of the playlist is an unlikely sequence of bytes or a likely sequence of bytes; and

responsive to determining that the overall perplexity score indicates that the sequence of bytes corresponding to the user-generated title of the playlist is a likely sequence of bytes, cause a recommendation system associated with the streaming service to recommend the playlist based on the overall perplexity score of the user-generated title.

8. The system of claim 7 , wherein the trained language model is trained using a plurality of training data that includes at least one of a plurality of pre-approved professionally curated playlist titles and a plurality of pre-approved professionally curated playlist descriptions.

9. The system of claim 7 , wherein the trained language model is trained using a plurality of training data that includes a plurality of pre-approved professionally curated playlist titles and wherein the trained language model is configured to generate the plurality of perplexity scores of the user-generated title of the playlist based on a negative log-likelihood that the sequence of bytes corresponding to the user-generated title of the playlist would appear in one of the plurality of pre-approved professionally curated playlist titles in the plurality of training data.

10. The system of claim 9 , wherein the hardware processor is further configured to:

responsive to determining that the overall perplexity score indicates that the sequence of bytes corresponding to the user-generated title is an unlikely sequence of bytes, cause the recommendation system associated with the streaming service to exclude the playlist from being recommended.

11. The system of claim 7 , wherein the hardware processor is further configured to determine that the user-generated title is based in a particular language from a plurality of languages.

12. The system of claim 7 , wherein the playlist comprises a plurality of music tracks.

13. A non-transitory computer-readable medium containing computer executable instructions that, when executed by a processor, cause the processor to:

detect a user-generated title associated with a playlist available on a streaming service;

generate a byte-level representation of the user-generated title based on the user-generated title of the playlist;

determine an embedded representation of the user-generated title based on the byte-level representation;

determine a plurality of perplexity scores that each correspond to an individual perplexity score of a portion of the embedded representation by inputting the embedded representation of the user-generated title into a trained language model, wherein the plurality of perplexity scores is an output of the trained language model;

determine an overall perplexity score of the user-generated title based on the plurality of perplexity scores, a number of bytes in the byte-level representation, and a position of each byte in the byte-level representation, wherein the overall perplexity score indicates whether a sequence of bytes corresponding to the user-generated title of the playlist is an unlikely sequence of bytes or a likely sequence of bytes; and

responsive to determining that the overall perplexity score indicates that the sequence of bytes corresponding to the user-generated title of the playlist is a likely sequence of bytes. cause a recommendation system associated with the streaming service to recommend the playlist based on the overall perplexity score of the user-generated title.

14. The non-transitory computer-readable medium of claim 13 , wherein the trained language model is trained using a plurality of training data that includes at least one of a plurality of pre-approved professionally curated playlist titles and a plurality of pre-approved professionally curated playlist descriptions.

15. The non-transitory computer-readable medium of claim 13 , wherein the trained language model is trained using a plurality of training data that includes a plurality of pre-approved professionally curated playlist titles and wherein the trained language model is configured to generate the plurality of perplexity scores of the user-generated title of the playlist based on a negative log-likelihood that the sequence of bytes corresponding to the user-generated title of the playlist would appear in one of the plurality of pre-approved professionally curated playlist titles in the plurality of training data.

16. The non-transitory computer-readable medium of claim 13 , wherein the execution of the instructions further causes the processor to:

responsive to determining that the overall perplexity score indicates that the sequence of bytes corresponding to the user-generated title is an unlikely sequence of bytes, cause the recommendation system associated with the streaming service to exclude the playlist from being recommended.

17. The non-transitory computer-readable medium of claim 13 , wherein execution of the instructions further causes the processor to determine that the user-generated title is based in a particular language from a plurality of languages.

18. The non-transitory computer-readable medium of claim 13 , wherein the playlist comprises a plurality of music tracks.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 30, 2022
From: GOODRICH, BEN; CHIPPALA, KUMAR
To: GOOGLE LLC
Reel/Frame 061923/0645 →
Continuity (1)
Related Publication 20240176816A1 · May 30, 2024
References Cited (23)
US 8914384B2 · Gates et al. · 2014 [cited by applicant]
US 9524487B1 · Yagnik et al. · 2016 [cited by applicant]
US 9563703B2 · Nijim et al. · 2017 [cited by applicant]
US 9589237B1 · Qamar · 2017 [cited by applicant]
US 10936653B2 · Levy et al. · 2021 [cited by applicant]
US 20200278997A1 · Lamere · 2020 [cited by examiner]
US 20200364299A1 · Niu · 2020 [cited by examiner]
US 20210374361A1 · Wick · 2021 [cited by examiner]
US 20220245362A1 · Nizar · 2022 [cited by examiner]
US 20230367968A1 · Eisenstadt · 2023 [cited by examiner]
US 20240028740A1 · Chan · 2024 [cited by examiner]
US 20240126924A1 · Pabolu · 2024 [cited by examiner]
Lim, Jong Yoon, et al. “Subsentence extraction from text using coverage-based deep learning language models.” Sensors 21.8 (2021): 2712. (Year: 2021). [cited by examiner]
Lau, “Cross-Entropy, Negative Log-Likelihood, and All That Jazz”, published in Toward Data Science, Mar. 8, 2022 (Year: 2022). [cited by examiner]
Lei, Lei. “Intelligent Recognition English Translation Model Based on Embedded Machine Learning and Improved GLR Algorithm.” Mobile Information Systems 2022 (Year: 2022). [cited by examiner]
Ruder, Sebastian, Ivan Vulić, and Anders Søgaard. “A survey of cross-lingual word embedding models.” Journal of Artificial Intelligence Research 65 (2019) (Year: 2019). [cited by examiner]
Backfried et al., “Method of expanding a vocabulary of a speech system”, EP 1074 973 A2, Feb. 7, 2001 (Year: 2001). [cited by examiner]
Prabhakaran Sethuraman, WO 2023140904 A1, PCT/US2022/048116, Jan. 21, 2022 (Year: 2022). [cited by examiner]
Baisa, Vit. Byte level language models. Diss. Ph. D. thesis, Masaryk University, 2016. (Year: 2016). [cited by examiner]
Gerz, Daniela, et al. “Language modeling for morphologically rich languages: Character-aware modeling for word-level prediction.” Transactions of the Association for Computational Linguistics 6 (2018): 451-465. (Year: 2… [cited by examiner]
Le Godais, Gaël, Tal Linzen, and Emmanuel Dupoux. “Comparing character-level neural language models using a lexical decision task.” Proceedings of the 15th Conference of the European Chapter of the Association for Compu… [cited by examiner]
Kim, Yoon, et al. “Character-aware neural language models.” Proceedings of the AAAI conference on artificial intelligence. vol. 30. No. 1. 2016. (Year: 2016). [cited by examiner]
Alvear, D., “Friendshipify: A Playlist Generator for Friends”, last updated May 27, 2020, pp. 1-14, available at: https://towardsdatascience.com/friendshipify-a-playlist-generator-for-friends-f79297f08b03. [cited by applicant]