IP Library › Granted Patent US 11,983,461
Granted Patent B2
US 11,983,461 · App. 17/211,321 · Granted May 14, 2024

Speech-based selection of augmented reality content for detected objects

Inventors: Joseph Timothy Fortier (Los Angeles, CA); Celia Nicole Mourkogiannis (Los Angeles, CA); Evan Spiegel (Los Angeles, CA); Kaveh Anvaripour (Santa Monica, CA)
Assignee: Snap Inc.
G06F3/167G06T11/00G06V10/454G06V10/764G06V20/10G06V20/20G06V20/64G06V40/161G06V40/168G06V40/174G10L15/08G10L15/22H04L51/046H04N23/60G06T2200/24G10L2015/088G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,983,461
App. No.
17/211,321
Filed
Mar 24, 2021
Granted
May 14, 2024
Kind
B2
Examiner
WANG, YI
Art Unit
2611
USPC
345/633
Abstract

Aspects of the present disclosure involve a system comprising a computer-readable storage medium storing a program and method for displaying augmented reality content. The program and method provide for causing, by a messaging application running on a device, a camera of the device to capture an image; receiving by the messaging application, speech input to select augmented reality content for display with the image; determining at least one keyword included in the speech input; determining that the at least one keyword indicates an object depicted in the image and an action to perform with respect to the object; identifying, from plural augmented reality content items, an augmented reality content item that corresponds to performing the action with respect to the object; and displaying the augmented reality content item with the image.

Claims (63)

1. A method, comprising:

causing, by a messaging application running on a device, a camera of the device to capture an image;

receiving by the messaging application, speech input to select augmented reality content for display with the image;

determining at least one keyword included in the speech input;

determining that the at least one keyword indicates an object depicted in the image and an action to perform with respect to the object;

determining first attributes of the object;

assigning weights to each of the first attributes of the object;

ranking plural augmented reality content items based on the assigned weights and on second attributes of the action;

selecting, based on the ranking, a highest-ranked augmented reality content item from among the plural augmented reality content items;

activating the highest-ranked augmented reality content item with respect to the image; and

displaying, in ranked order based on the ranking, an interface with user-selectable elements for activating remaining ones of the plural augmented reality content items.

2. The method of claim 1 , further comprising:

performing a scan of the image to identify multiple objects in the image; and

detecting, based on performing the scan, the object from among the multiple objects.

3. The method of claim 1 , wherein determining the at least one keyword comprises:

sending, to a speech recognition service, a request to perform speech recognition based on the speech input; and

receiving, from the speech recognition service and based on sending the request, the at least one keyword.

4. The method of claim 3 , wherein a first part of the speech input includes a trigger word, and

wherein the at least one keyword is based on a second part of the speech input that does not include the trigger word.

5. The method of claim 1 ,

wherein the interface is a carousel interface with a respective user-selectable icon for each of the plural augmented reality content items, and

wherein the carousel interface differentiates display of the icon for the augmented reality content item, relative to remaining icons, within the carousel interface.

6. The method of claim 1 , where the at least one keyword comprises a first keyword indicating the object and a second keyword indicating the action to perform with respect to the object.

7. The method of claim 1 , wherein the object is a math problem and the action corresponds to solving the math problem,

wherein the selecting comprises selecting, from among the plural augmented reality content items, the augmented reality content item that corresponds to displaying a solution to the math problem, and

wherein the displaying comprises displaying the solution to the math problem with the image.

8. The method of claim 7 , further comprising:

displaying a link which is user-selectable to navigate to a third-party application that provides additional information with respect to solving the math problem.

9. A device, comprising:

a processor; and

a memory storing instructions that, when executed by the processor, cause the processor to:

cause, by a messaging application running on a device, a camera of the device to capture an image;

receive by the messaging application, speech input to select augmented reality content for display with the image;

determine at least one keyword included in the speech input;

determine that the at least one keyword indicates an object depicted in the image and an action to perform with respect to the object;

determine first attributes of the object;

assign weights to each of the first attributes of the object;

rank plural augmented reality content items based on the assigned weights and on second attributes of the action;

select, based on the ranking, a highest-ranked augmented reality content item from among the plural augmented reality content items;

activate the highest-ranked augmented reality content item with respect to the image; and

display, in ranked order based on the ranking, an interface with user-selectable elements for activating remaining ones of the plural augmented reality content items.

10. The device of claim 9 , wherein the instructions further cause the processor to:

perform a scan of the image to identify multiple objects in the image; and

detect, based on performing the scan, the object from among the multiple objects.

11. The device of claim 9 , wherein determining the at least one keyword comprises:

sending, to a speech recognition service, a request to perform speech recognition based on the speech input; and

receiving, from the speech recognition service and based on sending the request, the at least one keyword.

12. The device of claim 11 , wherein a first part of the speech input includes a trigger word, and

wherein the at least one keyword is based on a second part of the speech input that does not include the trigger word.

13. The device of claim 9 , wherein

the interface is a carousel interface with a respective user-selectable icon for each of the plural augmented reality content items, and

wherein the carousel interface differentiates display of the icon for the augmented reality content item, relative to remaining icons, within the carousel interface.

14. A non-transitory computer-readable storage medium, the computer-readable storage medium including instructions that when executed by a computer, cause the computer to:

cause, by a messaging application running on a device, a camera of the device to capture an image;

receive by the messaging application, speech input to select augmented reality content for display with the image;

determine at least one keyword included in the speech input;

determine that the at least one keyword indicates an object depicted in the image and an action to perform with respect to the object;

determine first attributes of the object;

assign weights to each of the first attributes of the object;

rank plural augmented reality content items based on the assigned weights and on second attributes of the action;

select, based on the ranking, a highest-ranked augmented reality content item from among the plural augmented reality content items;

activate the highest-ranked augmented reality content item with respect to the image; and

display, in ranked order based on the ranking, an interface with user-selectable elements for activating remaining ones of the plural augmented reality content items.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 27, 2024
From: FORTIER, JOSEPH TIMOTHY; MOURKOGIANNIS, CELIA NICOLE; SPIEGEL, EVAN; ANVARIPOUR, KAVEH
To: SNAP INC.
Reel/Frame 066921/0066 →
Continuity (2)
Provisional Application 63000071 · Mar 26, 2020
Related Publication 20210304451A1 · Sep 30, 2021
Cited By (2)
US 12,211,504 US 12,481,367