IP Library Granted Patent US 8,914,284
Granted Patent B1
US 8,914,284 · App. 14/013,278 · Granted Dec 16, 2014

Methods and apparatus for conducting internet protocol telephony communication

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,914,284
App. No.
14/013,278
Granted
Dec 16, 2014
Kind
B1
Abstract

IP telephony communications are conducted by sending both data produced by a CODEC that represents received spoken audio input, and a textual representation of the spoken audio input. A receiving device utilizes the textual representation of the spoken audio input to help recreate the spoken audio input when a portion of the CODEC data is missing. The textual representation can be generated by a speech-to-text function. Alternatively, the textual representation can be a notation of extracted phonemes.

Claims (48)

1. A method of converting and transmitting audio information, comprising:

receiving spoken audio input;

converting the received spoken audio input into digital data that is representative of the received spoken audio input, wherein converting the received spoken audio input into digital data comprises creating a stream of digital data pockets, each digital data packet having a payload of data representative of a portion of the received spoken audio input;

generating a textual representation of the received spoken audio input;

inserting portions of the textual representation of the received spoken audio input into one or more headers of the digital data packets; and

transmitting corresponding portions of the digital data and the textual representation to a destination device at substantially the same time.

2. The method of claim 1 , wherein converting the received spoken audio input into digital data comprises using a CODEC to convert the received spoken audio input into digital data packets.

3. The method of claim 1 , wherein generating a textual representation of the received spoken audio input comprises performing a speech-to-text conversion of the received spoken audio input.

4. The method of claim 1 , wherein generating a textual representation of the received spoken audio input comprises:

extracting phonemes from the received spoken audio input; and

generating a textual representation of the extracted phonemes.

5. The method of claim 4 , wherein extracting phonemes from the received spoken audio input comprises:

determining the language of the received spoken audio input; and

extracting phonemes from the received spoken audio input based on the determined language.

6. The method of claim 1 , wherein the portion of the textual representation of the received spoken audio input that is inserted into the one or more headers of a digital data packet corresponds at least in part to a different portion of the received spoken audio input than the data in the payload of the digital data packet.

7. A system for converting and transmitting audio information, comprising:

means for receiving spoken audio input;

means for converting the received spoken audio input into digital data that is representative of the received spoken audio input, wherein the converting means creates a stream of digital data packets, each digital data packet having a payload of data representative of a portion of the received spoken audio input;

means for generating a textual representation of the received spoken audio input; and

means for transmitting corresponding portions of the digital data and the textual representation to a destination device at substantially the same time, wherein the transmitting means inserts portions of the textual representation of the received spoken audio input generated by the textual representation generating means into one or more headers of the digital data packets.

8. A system for converting and transmitting audio information, comprising:

an audio input receiving unit that receives spoken audio input;

a conversion unit that converts the received spoken audio input into digital data that is representative of the received spoken audio input, wherein the conversion unit creates a stream of digital data packets, each digital data packet having a payload of data representative of a portion of the received spoken audio input;

a textual representation conversion unit that generates a textual representation of the received spoken audio input; and

a transmission unit that transmits corresponding portions of the digital data and the textual representation to a destination device at substantially the same time, wherein the transmission unit inserts portions of the textual representation of the received spoken audio input generated by the textual representation conversion unit into one or more headers of the digital data packets.

9. The system of claim 8 , wherein the conversion unit uses a CODEC to convert the received spoken audio input into digital data packets.

10. The system of claim 8 , wherein the textual representation conversion unit performs a speech-to-text conversion of the received spoken audio input.

11. The system of claim 8 , wherein the textual representation conversion unit extracts phonemes from the received spoken audio input, and generates a textual representation of the extracted phonemes.

12. The system of claim 11 , wherein the textual representation conversion unit extracts phonemes from the received spoken audio input by determining the language of the received spoken audio input, and then extracting phonemes from the received spoken audio input based on the determined language.

13. The system of claim 8 , wherein the portion of the textual representation of the received spoken audio input that is inserted into the one or more headers of a digital data packet corresponds, at least in part, to a different portion of the received spoken audio input than the data in the payload of the digital data packet.

14. A method of converting and transmitting audio information, comprising:

receiving spoken audio input;

converting the received spoken audio input into digital data that is representative of the received spoken audio input, wherein converting the received spoken audio input into digital data comprises creating a first stream of digital data packets, each digital data packet in the first stream having a payload of data representative of a portion of the received spoken audio input;

generating a textual representation of the received spoken audio input, wherein generating a textual representation of the received spoken input comprises creating a second stream of digital data packets, each digital data packet in the second stream having a payload of data that is representative of a portion of the textual representation of the received spoken audio input; and

transmitting corresponding portions of the digital data and the textual representation to a destination device at substantially the same time, wherein the transmitting step comprises sending the first and second streams of digital data packets to the destination device.

15. The method of claim 14 , wherein generating a textual representation of the received spoken audio input comprises:

extracting phonemes from the received spoken audio input; and

generating a textual representation of the extracted phonemes.

16. The method of claim 15 , wherein extracting phonemes from the received spoken audio input comprises:

determining the language of the received spoken audio input; and

extracting phonemes from the received spoken audio input based on the determined language.

17. A system for converting and transmitting audio information, comprising:

an audio input receiving unit that receives spoken audio input;

a conversion unit that converts the received spoken audio input into digital data that is representative of the received spoken audio input, wherein the conversion unit creates a first stream of digital data packets, each digital data packet in the first stream having a payload of data representative of a portion of the received spoken audio input;

a textual representation conversion unit that generates a textual representation of the received spoken audio input, wherein the textual representation conversion unit creates a second stream of digital data packets, each digital data packet in the second stream having a payload of data that is representative of a portion of the textual representation of the received spoken audio input; and

a transmission unit that transmits corresponding portions of the digital data and the textual representation to a destination device at substantially the same time, wherein the transmitting unit sends the first and second streams of digital data packets to the destination device.

18. The system of claim 17 , wherein the textual representation conversion unit extracts phonemes from the received spoken audio input, and generates a textual representation of the extracted phonemes.

19. The system of claim 18 , wherein the textual representation conversion unit extracts phonemes from the received spoken audio input by determining the language of the received spoken audio input, and then extracting phonemes from the received spoken audio input based on the determined language.

Assignments (8)
RELEASE OF SECURITY INTEREST Recorded Jul 28, 2022
From: JPMORGAN CHASE BANK, N.A.
To: VONAGE AMERICA INC.; VONAGE HOLDINGS CORP.; VONAGE BUSINESS INC.; NEXMO INC.; TOKBOX, INC.
Reel/Frame 061002/0340 →
SECURITY INTEREST Recorded Nov 12, 2018
From: VONAGE BUSINESS INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 047502/0432 →
CORRECTIVE ASSIGNMENT TO CORRECT THE LIST BY DELETING 13831728 13831785 14291602 13680382 14827548 14752086 13680067 14169385 14473289 14194220 14194438 14317743 PREVIOUSLY RECORDED ON REEL 038328 FRAME 501. ASSIGNOR(S) HEREBY CONFIRMS THE SALE, ASSIGNMENT, TRANSFER AND CONVEYANCE OF REMAINING PROPERTIES. Recorded Oct 28, 2016
From: VONAGE NETWORK LLC
To: VONAGE BUSINESS INC.
Reel/Frame 040540/0702 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 1, 2016
From: VONAGE NETWORK LLC
To: VONAGE BUSINESS INC.
Reel/Frame 038328/0501 →
CORRECTIVE ASSIGNMENT TO CORRECT THE PATENT APPLICATION NUMBER 13966486 PREVIOUSLY RECORDED ON REEL 033545 FRAME 0424. ASSIGNOR(S) HEREBY CONFIRMS THE SECURITY INTEREST. Recorded Jan 21, 2016
From: VONAGE HOLDINGS CORP.; VONAGE NETWORK LLC; VONAGE BUSINESS SOLUTIONS INC.; VONAGE AMERICA INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 037570/0203 →
SECURITY INTEREST Recorded Jul 29, 2015
From: VONAGE HOLDINGS CORP.; VONAGE AMERICA INC.; VONAGE BUSINESS SOLUTIONS, INC.; VONAGE NETWORK LLC
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 036205/0485 →
SECURITY INTEREST Recorded Aug 14, 2014
From: VONAGE HOLDINGS CORP.; VONAGE NETWORK LLC; VONAGE BUSINESS SOLUTIONS INC.; VONAGE AMERICA INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 033545/0424 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 29, 2013
From: STERMAN, BARUCH; EFRATI, TZAHI; BIANCO, ITAY; MACHLIN, SAGIE; MINTZ, IDO
To: VONAGE NETWORK, LLC
Reel/Frame 031108/0787 →