IP Library Granted Patent US 8,635,243
Granted Patent B2
US 8,635,243 · App. 12/870,257 · Granted Jan 21, 2014

Sending a communications header with voice recording to send metadata for use in speech recognition, formatting, and search mobile search application

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,635,243
App. No.
12/870,257
Granted
Jan 21, 2014
Kind
B2
Abstract

In embodiments of the present invention improved capabilities are described for sending a communications header with the voice recording to send metadata for use in speech recognition, formatting, and search in searching for web content on a mobile communication facility comprising capturing speech presented by a user using a resident capture facility on the mobile communication facility; transmitting a communications header to a speech recognition facility from the mobile communication facility through a wireless communications facility, wherein the communications header includes at least one of device name, network type, audio source, display parameters for the wireless communications facility, geographic location, and phone number information; transmitting at least a portion of the captured speech as data through the wireless communication facility to a speech recognition facility; generating speech-to-text results utilizing the speech recognition facility based at least in part on the information relating to the captured speech and the communications header; and transmitting text from the speech-to-text results along with URL usage information configured to enable a user to conduct a search on the mobile communication facility.

Claims (20)

1. A method of searching for web content on a mobile communication facility comprising:

capturing, by the mobile communication facility, having one or more processors, speech presented by a user using a resident capture facility on the mobile communication facility;

transmitting, by the mobile communication facility, a communications header to a speech recognition facility from the mobile communication facility through a wireless communications facility, wherein the communications header includes at least one of device name, network type, audio source, display parameters for the wireless communications facility, geographic location, or phone number information;

transmitting, by the mobile communication facility, at least a portion of the captured speech as data through the wireless communication facility to a speech recognition facility;

receiving, by the mobile communication facility, text from the speech recognition facility from speech-to-text results generated by the speech recognition facility based at least in part on the data relating to the at least a portion of the captured speech and based on the communications header, the text including URL usage information configured to enable a user to conduct the search on the mobile communication facility, wherein the Uniform Resource Locators (URL) usage information is included for a plurality of distinct web-based search facilities; and

searching for web content on the mobile communication facility.

2. The method of claim 1 , wherein the URL usage information is associated with a plurality of base URLs and includes at least one formatting rule specifying how search text is combined with a base URL to form a search URL that enables a search based on the text using the mobile communications facility.

3. The method of claim 1 , wherein the mobile communication facility comprises a resident search engine for searching for the web content.

4. The method of claim 1 , wherein the step of receiving comprises receiving the text from the speech recognition facility into a text field on the mobile communications facility.

5. The method of claim 1 , further comprising loading the loading the text into a search application resident on the mobile communications facility.

6. A method of searching for web content on a mobile communication facility comprising:

capturing, by the mobile communication facility, having one or more processors, speech presented by a user using a resident capture facility on the mobile communication facility;

transmitting, by the mobile communication facility, a communications header to a speech recognition facility from the mobile communication facility through a wireless communications facility, wherein the communications header includes metadata with at least one of device name, a network type, an audio source, display parameters for the wireless communications facility, a geographic location, or phone number information;

transmitting, by the mobile communication facility, at least a portion of the captured speech as data through the wireless communication facility to a speech recognition facility;

receiving, by the mobile communication facility, text from the speech recognition facility from speech-to-text results generated by the speech recognition facility based at least in part on information relating to the captured speech and the communications header, the text including information configured to enable a user to conduct a search on the mobile communication facility, wherein the speech-to-text results are generated using at least one statistical language model selected from a set of language models based at least in part on the information relating to the captured speech and the communications header, wherein the received text includes URL usage information configured to enable the user to conduct the search on the mobile communication facility; and

loading the text into a search application resident on the mobile communications facility to search for the web content.

7. The method of claim 6 , wherein the speech-to-text results are generated using client state information relating to the captured speech.

8. The method of claim 7 , wherein the client state information comprises information selected from the group consisting of a user identification, a usage history of the search application on the mobile communications device, and a location of the mobile communications device.

9. The method of claim 6 , further comprising deciding whether the at least one selected statistical language model provides insufficient recognition output for the speech-to-text results and repeating the step of receiving text from speech-to-test results, wherein the speech-to-text results are generated using at least one additional language model from the first set of language models for a subsequent step of receiving text.

10. The method of claim 6 , further comprising receiving results of the search on the mobile communication facility.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 13, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065552/0934 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 27, 2019
From: RESEARCH IN MOTION LIMITED
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 050509/0308 →
SECURITY AGREEMENT Recorded Dec 12, 2011
From: VLINGO CORPORATION
To: RESEARCH IN MOTION LIMITED
Reel/Frame 027362/0898 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 14, 2010
From: PHILLIPS, MICHAEL S.; NGUYEN, JOHN N.
To: VLINGO CORPORATION
Reel/Frame 025138/0911 →