IP Library Granted Patent US 9,767,802
Granted Patent B2
US 9,767,802 · App. 14/570,599 · Granted Sep 19, 2017

Methods and apparatus for conducting internet protocol telephony communications

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,767,802
App. No.
14/570,599
Granted
Sep 19, 2017
Kind
B2
Abstract

IP telephony communications are conducted by sending both audio data produced by a CODEC that represents received spoken audio input, and a textual representation of the spoken audio input. A receiving device utilizes the textual representation of the spoken audio input to help recreate the spoken audio input when a portion of the CODEC data is missing. The textual representation can be generated by a speech-to-text function. Alternatively, the textual representation can be a notation of extracted phonemes.

Claims (42)

1. A method of converting and transmitting audio information, comprising:

receiving spoken audio input;

converting the received spoken audio input into audio digital data that is representative of the received spoken audio input;

generating a stream of audio digital data packets that contain the audio digital data;

generating a textual representation of the received spoken audio input;

generating a stream of textual digital data packets that contain the generated textual representation of the received spoken audio input; and

transmitting the stream of audio digital data packets and the stream of textual digital data packets to a destination device;

wherein at least one of the audio digital data packets and/or the textual digital data packets include information that indicates which audio digital data packets and textual digital data packets contain data relating to the same portions of the received spoken audio input.

2. The method of claim 1 , wherein the each textual digital data packet includes sequence number information for at least one audio digital data packet that contains audio digital data that corresponds to the same portion of the received audio input as the textual digital data that is loaded in the textual digital data packet.

3. The method of claim 1 , wherein generating a textual representation of the received spoken audio input comprises:

extracting phonemes from the received spoken audio input; and

generating a textual representation of the extracted phonemes.

4. The method of claim 3 , wherein extracting phonemes from the received spoken audio input comprises:

determining the language of the received spoken audio input; and

extracting phonemes from the received spoken audio input based on the determined language.

5. The method of claim 1 , wherein converting the received spoken audio input into digital data comprises using a CODEC to convert the received spoken audio input into audio digital data.

6. The method of claim 1 , wherein generating a textual representation of the received spoken audio input comprises performing a speech-to-text conversion of the received spoken audio input.

7. The method of claim 1 , wherein each textual digital data packet includes sequence numbers assigned to the audio digital data packets that contain audio digital data that corresponds to the same portion of the received audio input as the textual digital data that is loaded in the textual digital data packet.

8. The method of claim 1 , wherein each audio digital data packet includes a sequence number assigned to at least one of the textual digital data packets that contains textual digital data that corresponds to the same portion of the received audio input as the audio digital data that is loaded in the audio digital data packet.

9. A system for converting and transmitting audio information, comprising:

means for receiving spoken audio input;

means for converting the received spoken audio input into audio digital data that is representative of the received spoken audio input;

means for generating a stream of audio digital data packets that contain the audio digital data;

means for generating a textual representation of the received spoken audio input;

means for generating a stream of textual digital data packets that contain the generated textual representation of the received spoken audio input; and

means for transmitting the stream of audio digital data packets and the stream of textual digital data packets to a destination device;

wherein at least one of the audio digital data packets and/or the textual digital data packets include information that indicates which audio digital data packets and digital data packets contain data relating to the same portions of the received spoken audio input.

10. A system for converting and transmitting audio information, comprising:

an audio input receiving unit that receives spoken audio input;

a first conversion unit that converts the received spoken audio input into audio digital data that is representative of the received spoken audio input;

an audio digital data packet generation unit that generates a stream of audio digital data packets that contain the audio digital data;

a textual representation conversion unit that generates a textual representation of the received spoken audio input;

a textual digital data packet generation unit that generates a stream of textual digital data packets that contain the generated textual representation of the received spoken audio input; and

a transmission unit that transmits the stream of audio digital data packets and the stream of textual digital data packets to a destination device;

wherein at least one of the audio digital data packets and/or the textual digital data packets include information that indicates which audio digital data packets and digital data packets contain data relating to the same portions of the received spoken audio input.

11. The system of claim 10 , wherein the each textual digital data packet includes sequence number information for at least one audio digital data packet that contains audio digital data that corresponds to the same portion of the received audio input as the textual digital data that is loaded in the textual digital data packet.

12. The system of claim 10 , wherein the textual representation conversion unit extracts phonemes from the received spoken audio input and generates a textual representation of the extracted phonemes.

13. The system of claim 12 , wherein the textual representation conversion unit determines the language of the received spoken audio input, and extracts phonemes from the received spoken audio input based on the determined language.

14. The system of claim 10 , wherein first conversion unit uses a CODEC to convert the received spoken audio input into audio digital data.

15. The system of claim 10 , wherein the textual representation conversion unit performs a speech-to-text conversion of the received spoken audio input.

16. The system of claim 10 , wherein the textual digital data packet generation unit inserts into each textual digital data packet sequence numbers assigned to the audio digital data packets that contain audio digital data that corresponds to the same portion of the received audio input as the textual digital data that is loaded in the textual digital data packet.

17. The system of claim 10 , wherein the audio digital data packet generation unit inserts into each audio digital data packet a sequence number assigned to at least one of the textual digital data packets that contains textual digital data that corresponds to the same portion of the received audio input as the audio digital data that is loaded in the audio digital data packet.

Assignments (3)
RELEASE OF SECURITY INTEREST Recorded Jul 28, 2022
From: JPMORGAN CHASE BANK, N.A.
To: VONAGE AMERICA INC.; VONAGE HOLDINGS CORP.; VONAGE BUSINESS INC.; NEXMO INC.; TOKBOX, INC.
Reel/Frame 061002/0340 →
SECURITY INTEREST Recorded Nov 12, 2018
From: VONAGE BUSINESS INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 047502/0432 →
CORRECTIVE ASSIGNMENT TO CORRECT THE LIST BY DELETING 13831728 13831785 14291602 13680382 14827548 14752086 13680067 14169385 14473289 14194220 14194438 14317743 PREVIOUSLY RECORDED ON REEL 038328 FRAME 501. ASSIGNOR(S) HEREBY CONFIRMS THE SALE, ASSIGNMENT, TRANSFER AND CONVEYANCE OF REMAINING PROPERTIES. Recorded Oct 28, 2016
From: VONAGE NETWORK LLC
To: VONAGE BUSINESS INC.
Reel/Frame 040540/0702 →