IP Library Granted Patent US 7,050,978
Granted Patent B2
US 7,050,978 · App. 10/132,980 · Granted May 23, 2006

System and method of providing evaluation feedback to a speaker while giving a real-time oral presentation

Assignee: Hewlett-Packard Development Company, L.P.
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,050,978
App. No.
10/132,980
Filed
Apr 26, 2002
Granted
May 23, 2006
Kind
B2
Art Unit
2626
USPC
704/271
Abstract

A system for and method of providing feedback information relating to characteristics of the oral presentation to a speaker while giving a real-time oral presentation by analyzing representations of the audio signal corresponding to the oral presentation. The feedback information can then be provided to the speaker during the real-time presentation to assist the speaker and improve the oral presentation.

Claims (43)

1. A system comprising:

signal processor for processing an audio signal corresponding to and during a real-time oral presentation to generate at least one representation of the audio signal including at least an energy function representation and a zero-crossing rate function representation;

analyzer for analyzing at least the one representation to generate at least one characterizing indicator corresponding to the oral presentation;

output device for providing feedback information characterizing the oral presentation in response to the at least one characterizing indicator.

2. The system as described in claim 1 wherein the feedback information is provided during the real-time oral presentation on a real-time basis.

3. The system as described in claim 1 wherein the feedback information is provided after the real-time oral presentation.

4. The system as described in claim 1 wherein the output device stores feedback information in the form of stored phrases.

5. The system as described in claim 1 further comprising a volume analyzer for detecting percentage of energy peaks in a given interval of the energy function representation having a magnitude greater or less than preselected threshold values and comparing to a preselected percentage.

6. The system as described in claim 1 further comprising a pace analyzer for detecting the number of peaks within a given interval of the energy function representation to identify number of syllables spoken in the interval and comparing the identified number of syllables spoken to a preselected range to identify when the oral presentation does not conform to a preselected pace.

7. The system as described in claim 1 further comprising a filler word analyzer for detecting flat intervals within the zero-crossing rate function representation corresponding to filler words in the oral presentation.

8. The system as described in claim 1 further comprising a filler word analyzer for detecting intervals without significant troughs within the energy function representation corresponding to filler words in the oral presentation.

9. The system as described in claim 1 further comprising a pause analyzer for detecting intervals in the range of none to very low energy within the energy function representation corresponding to pauses in the oral presentation.

10. The system as described in claim 1 further comprising a tone analyzer for detecting variances of amplitude of energy peaks within the energy function representation to determine a tone variation value of the oral presentation and comparing the tone variation to preselected tone variation threshold values.

11. The system as described in claim 1 further comprising an oral presentation time analyzer comprising:

a speech recognizer for identifying a key word in the audio signal;

a means for linking the key word to a slide associated with the oral presentation and predetermined time information associated with the slide.

12. A method comprising:

processing an audio signal corresponding to and during a real-time oral presentation so as to generate at least one representation of the audio signal including at least an energy function representation and a zero-crossing rate function representation;

analyzing at least the one representation to obtain at least one characterizing indicator corresponding to the oral presentation;

determining feedback information from the at least one characterization indicator;

providing the feedback information characterizing the real-time oral presentation.

13. The method as described in claim 12 comprising providing the feedback information during the oral presentation on a real-time basis.

14. The method as described in claim 12 comprising providing the feedback information after the oral presentation.

15. The method as described in claim 12 further comprising detecting the percentage of energy peaks in a given interval of the energy function representation having a magnitude greater or less than preselected threshold values and comparing to a preselected percentage.

16. The method as described in claim 12 further comprising detecting the number of peaks within a given interval of the energy function representation to identify number of syllables spoken in the interval and comparing the identified number of syllables spoken to a preselected range to identify when the oral presentation does not conform to a preselected pace.

17. The method as described in claim 12 comprising detecting flat intervals within the zero-crossing rate function representation corresponding to filler words in the oral presentation.

18. The method as described in claim 12 comprising detecting intervals without significant troughs within the energy function representation corresponding to filler words in the oral presentation.

19. The method as described in claim 12 comprising detecting intervals in the range of none to very low energy within the energy function representation corresponding to pauses in the oral presentation.

20. The method as described in claim 12 comprising detecting variances of amplitude of energy peaks within the energy function representation to determine a tone variation value of the oral presentation and comparing the tone variation to preselected tone variation threshold values.

21. The method as described in claim 12 comprising analyzing oral presentation time characteristics by identifying key words in the audio signal and linking the key words to a slide associated with the oral presentation and a predetermined time information associated with the slide.

22. A computer readable medium for causing a processor in a computer system to perform processing instructions comprising:

processing an audio signal corresponding to and during a real-time oral presentation so as to generate at least one representation of the audio signal including at least an energy function representation and a zero-crossing rate function representation;

analyzing at least the one representation to obtain at least one characterizing indicator corresponding to the oral presentation;

determining feedback information from the at least one characterization indicator;

providing the feedback information characterizing the real-time oral presentation.

23. A system comprising a processor for:

processing an audio signal corresponding to and during a real-time oral presentation so as to generate at least one representation of the audio signal including at least an energy function representation and a zero-crossing rate function representation;

analyzing at least the one representation to obtain at least one characterizing indicator corresponding to the oral presentation;

determining feedback information from the at least one characterization indicator;

providing the feedback information characterizing the real-time oral presentation.

24. The system as described in claim 23 further comprising an output device for providing the feedback information by one of an audio signal, a visual signal, and a combination of an audio and visual signal.

25. The system as described in claim 24 wherein the output device is one of an earphone, display screen, and a printer.

26. The system as described in claim 23 wherein the feedback information is one of command phrases and evaluation phrases.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 18, 2003
From: HEWLETT-PACKARD COMPANY
To: HEWLETT-PACKARD DEVELOPMENT COMPANY, L.P.
Reel/Frame 013776/0928 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 21, 2002
From: ZHANG, TONG; SILVERSTEIN, D. AMNON
To: HEWLETT-PACKARD COMPANY
Reel/Frame 013397/0497 →
Continuity (1)
Related Publication 20030202007A1 · Oct 30, 2003