IP Library Granted Patent US 9,484,046
Granted Patent B2
US 9,484,046 · App. 14/460,719 · Granted Nov 1, 2016

Smartphone-based methods and systems

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,484,046
App. No.
14/460,719
Granted
Nov 1, 2016
Kind
B2
Abstract

Arrangements involving portable devices (e.g., smartphones and tablet computers) are disclosed. One arrangement enables a content creator to select software with which that creator's content should be rendered—assuring continuity between artistic intention and delivery. Another utilizes a device camera to identify nearby subjects, and take actions based thereon. Others rely on near field chip (RFID) identification of objects, or on identification of audio streams (e.g., music, voice). Some technologies concern improvements to the user interfaces associated with such devices. For example, some arrangements enable discovery of both audio and visual content, without any user requirement to switch modes. Other technologies involve use of these devices in connection with shopping, text entry, and vision-based discovery. Still other improvements are architectural in nature, e.g., relating to evidence-based state machines, and blackboard systems. Yet other technologies concern computational photography. A great variety of other features and arrangements are also detailed.

Claims (38)

1. A method comprising the acts:

receiving audio data using a microphone of a user's portable device;

receiving image data using a camera of said device;

recognition-processing both the received audio and image data—without a user being required to operate a user interface control to switch between an audio recognition mode and an image recognition mode, said recognition-processing being performed by a hardware processor configured to perform such act;

presenting a graphical user interface on the screen of the portable device, and displaying the image data received from the camera in a viewfinder region of said graphical user interface;

also presenting, in said graphical user interface, a stack of tiles, each of which corresponds to an item of recognized content, said stack including a first tile corresponding to a first item of recognized audio content, and a second tile corresponding to a second item of recognized visual content, wherein said user interface similarly represents items of recognized audio and visual content by tiles corresponding thereto;

said stack of tiles growing in size on said screen as successive items of content are recognized, until a first dimension is reached, after which older tiles disappear off the screen, wherein said first dimension assures that the viewfinder region is preserved for presentation of the image data from the camera;

each of the tiles in the stack having a payoff associated therewith, the user interface requiring user interaction with the first tile to initiate a first payoff corresponding to said item of recognized audio content, but not requiring user interaction with the second tile to initiate a second payoff corresponding to said item of recognized visual content, said second payoff instead being initiated automatically;

wherein said user interface, which similarly represents items of recognized audio and visual content by tiles corresponding thereto, differently treats initiations of payoffs for said items.

2. The method of claim 1 that includes storing entries in a history data structure, said entries identifying both recognized audio content and recognized visual content.

3. The method of claim 1 that includes, in response to user input, recalling older tiles that have disappeared off the screen, the tiles occupying the viewfinder region that formerly was preserved for presentation of the image data.

4. The method of claim 1 that further includes:

in response to user action, recalling back to the screen a third tile corresponding to a third item of recognized content that has disappeared off the screen;

wherein the third tile comprises a first image graphic when originally presented on the screen, before disappearing off the screen, and comprises a second image graphic, different than said first image graphic, when recalled back to the screen.

5. The method of claim 1 that further includes:

at a first time, inferring from context data that the user is interested in engaging in visual discovery, and positioning tiles from said recognition events to preserve said viewfinder region of the screen; and

at a second time, inferring from context data that the user is not interested in engaging in visual discovery, and positioning tiles from said recognition events over said viewfinder region of the screen.

6. A smartphone comprising a processor, a memory, a screen, a microphone and a camera, the memory containing instructions configuring the smartphone to perform acts including:

producing audio data using the microphone;

producing image data using the camera;

recognition-processing both the audio and image data—without a user being required to operate a user interface control to switch between an audio recognition mode and an image recognition mode;

presenting a graphical user interface on the screen, and displaying image data from the camera in a viewfinder region of said graphical user interface;

also presenting, in said graphical user interface, a stack of tiles, each of which corresponds to an item of recognized content, said stack including a first tile corresponding to a first item of recognized audio content, and a second tile corresponding to a second item of recognized visual content, said stack of tiles thereby similarly serving to represent instances of both audio and visual content recognition;

said stack of tiles growing in size on said screen as successive items of content are recognized, until a first dimension is reached, after which older tiles disappear off the screen, wherein said first dimension assures that the viewfinder region is preserved for presentation of the image data from the camera;

each of the tiles in the stack having a payoff associated therewith, the user interface requiring user interaction with the first tile to initiate a first payoff corresponding to said first item of recognized audio content, but not requiring user interaction with the second tile to initiate a second payoff corresponding to said second item of recognized visual content, said second payoff instead being initiated automatically;

wherein said user interface, which similarly represents items of recognized audio and visual content by tiles corresponding thereto, differently treats initiations of payoffs for said items.

7. The smartphone of claim 6 that further includes:

means for inferring whether or not the user is interested in engaging in visual discovery; and

wherein said instructions further configure the smartphone to:

at a first time, when said means infers that the user is interested in engaging in visual discovery, position tiles from said recognition events to preserve said viewfinder region of the screen; and

at a second time, when said means infers that the user is not interested in engaging in visual discovery, position tiles from said recognition events over said viewfinder region of the screen.

8. A non-transitory computer readable medium containing instructions for configuring a camera-equipped smartphone to perform acts including:

recognition-processing both received audio and image data—without a user being required to operate a user interface control to switch between an audio recognition mode and an image recognition mode;

presenting a graphical user interface on a screen of said smartphone, and displaying image data from the camera in a viewfinder region of said graphical user interface;

also presenting, in said graphical user interface, a stack of tiles, each of which corresponds to an item of recognized content, said stack including a first tile corresponding to a first item of recognized audio content, and a second tile corresponding to a second item of recognized visual content, wherein said user interface similarly represents items of recognized audio and visual content by tiles corresponding thereto;

said stack of tiles growing in size on said screen as successive items of content are recognized, until a first dimension is reached, after which older tiles disappear off the screen, wherein said first dimension assures that the viewfinder region is preserved for presentation of the image data from the camera;

each of the tiles in the stack having a payoff associated therewith, the user interface requiring user interaction with the first tile to initiate a first payoff corresponding to said first item of recognized audio content, but not requiring user interaction with the second tile to initiate a second payoff corresponding to said second item of recognized visual content, said second payoff instead being initiated automatically;

wherein said user interface, which similarly represents items of recognized audio and visual content by tiles corresponding thereto, differently treats initiations of payoffs for said items.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 29, 2016
From: KNUDSON, EDWARD B.; RHOADS, GEOFFREY B.; CORNABY, COLIN P.; SINCLAIR, EOIN C.; ROGERS, ELIOT
To: DIGIMARC CORPORATION
Reel/Frame 039296/0143 →