IP Library › Granted Patent US 12,033,635
Granted Patent B2
US 12,033,635 · App. 18/131,895 · Granted Jul 9, 2024

Image display apparatus and method of controlling the same

Inventors: Dae Gyu Bae (Suwon-si, KR); Tae Hwan Cha (Yongin-si, KR); Ho Jeong You (Suwon-si, KR)
Assignee: SAMSUNG ELECTRONICS CO., LTD.
G10L15/22G10L15/08G10L15/26G10L17/22G10L21/00H03G3/02H03G3/3005H04N21/42203H04N21/42204H04N21/4221H04N21/42222H04N21/4312H04N21/4394H04N21/4396G10L2015/223H04N21/42206
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,033,635
App. No.
18/131,895
Granted
Jul 9, 2024
Kind
B2
Abstract

Provided are an image display apparatus and a method of controlling the same. The image display apparatus enabling voice recognition includes: a first voice inputter which receives a user-side audio signal; an audio outputter which outputs an audio signal processed by the image display apparatus; a first voice recognizer which recognizes the user-side audio signal received through the first voice inputter; and a controller which decreases a volume of the audio signal output through the audio outputter to a predetermined level if a voice recognition start command is received.

Claims (82)

1. An image display apparatus comprising:

a first audio receiver;

a display to display a content;

a first communicator to receive a control signal from a remote controller; and

a controller configured to:

in response to receiving a first audio input for starting a speech recognition from the first audio receiver, start the speech recognition, and

in response to a second audio input being received via the first audio receiver after the speech recognition is started, perform an operation according to the second audio input, and control the display to display a result of the operation,

wherein the remote controller is configured to control the image display apparatus and includes:

a button for controlling the image display apparatus to start the speech recognition,

a second audio receiver to receive a third audio input wherein the third audio input is received as a user voice via the second audio receiver of the remote controller, and

a second communicator to communicate with the image display apparatus and to transmit the control signal to the image display apparatus,

wherein in response to the button of the remote controller being pressed while the image display apparatus waits to receive a user voice via the first audio receiver after the speech recognition is started by the first audio input to the first audio receiver of the image display apparatus, the controller is further configured to end the speech recognition started by the first audio input.

2. The image display apparatus according to claim 1 , wherein the controller is configured to, in response to receiving the first audio input for starting the speech recognition from the first audio receiver, initialize a standby time, and

in response to no audio input for speech recognition being received via the first audio receiver until the standby time reaches a first reference time after the start of the speech recognition, end the speech recognition started by the first audio input.

3. The image display apparatus according to claim 2 , wherein the controller is configured to, in response to the second audio input for speech recognition being received via the first audio receiver before the standby time reaches the first reference time after the start of the speech recognition, initialize the standby time.

4. The image display apparatus according to claim 1 , further comprising an audio output unit to output an audio signal,

wherein the controller is configured to, in response to receiving the first audio input for starting the speech recognition, decrease a volume of the audio signal output through the audio output unit to a predetermined level.

5. The image display apparatus according to claim 4 , wherein the predetermined level is not zero level.

6. The image display apparatus according to claim 4 , wherein the controller is configured to,

in response to receiving the first audio input for starting the speech recognition from the first audio receiver, initialize a standby time, and

in response to no audio input for speech recognition being received via the first audio receiver until the standby time reaches the first reference time after the start of the speech recognition, restore the volume of the audio signal output through the audio output unit to a volume level set before the speech recognition starts.

7. The image display apparatus according to claim 1 , further comprising an audio output unit to output an audio signal, wherein the controller is configured to:

in response to receiving the first audio input for starting the speech recognition from the first audio receiver, decrease a volume of the audio signal output through the audio output unit to a predetermined level and initialize a standby time,

after displaying the result of the operation according to the second audio input on the display, re-start the speech recognition for receiving a fourth audio input for speech recognition associated with the result and initialize the standby time,

in response to no audio input for speech recognition being received via the first receiver until the standby time reaches a second reference time after displaying the result on the display, end the speech recognition and restore the volume of the audio signal output through the audio output unit to a volume level set before the speech recognition starts.

8. The image display apparatus according to claim 7 , wherein the second reference time is different from the first reference time.

9. The image display apparatus according to claim 7 , wherein the second reference time is longer than the first reference time.

10. The image display apparatus according to claim 1 , further comprising a background sound canceller to cancel background sound other than the user voice from the first audio receiver.

11. A non-transitory computer-readable medium storing computer-executable instructions when executed by at least one processor of an image display apparatus, cause the image display apparatus to perform:

in response to receiving a first audio input for starting a speech recognition from a first audio receiver of the image display apparatus, starting the speech recognition,

in response to a second audio input being received via the first audio receiver after the speech recognition is started, performing an operation according to the second audio input, and controlling a display of the image display apparatus to display a result of the operation, and

in response to a button of a remote controller being pressed while the image display apparatus waits to receive a user voice via the first audio receiver after the speech recognition is started by the first audio input to the first audio receiver of the image display apparatus, ending the speech recognition started by the first audio input,

wherein the remote controller is configured to control the image display apparatus, and the button of the remote controller is configured to control the image display apparatus to start the speech recognition,

wherein the remote controller includes:

a second audio receiver to receive a third audio input wherein the third audio input is received as a user voice via the second audio receiver of the remote controller, and

a communicator to communicate with the image display apparatus and to transmit a control signal to the image display apparatus.

12. The non-transitory computer-readable medium according to claim 11 , wherein when executed by the at least one processor of the image display apparatus, the instructions cause the image display apparatus to further perform:

in response to receiving the first audio input for starting the speech recognition from the first audio receiver, initializing a standby time, and

in response to no audio input for speech recognition being received via the first audio receiver until the standby time reaches a first reference time after the start of the speech recognition, ending the speech recognition started by the first audio input.

13. The non-transitory computer-readable medium according to claim 12 , wherein when executed by the at least one processor of the image display apparatus, the instructions cause the image display apparatus to further perform:

in response to the second audio input for speech recognition being received via the first audio receiver before the standby time reaches the first reference time after the start of the speech recognition, initializing the standby time.

14. The non-transitory computer-readable medium according to claim 11 , wherein when executed by the at least one processor of the image display apparatus, the instructions cause the image display apparatus to further perform:

in response to receiving the first audio input for starting the speech recognition, decreasing a volume of an audio signal output through an audio output unit of the image display apparatus to a predetermined level.

15. The non-transitory computer-readable medium according to claim 14 , wherein the predetermined level is not zero level.

16. The non-transitory computer-readable medium according to claim 14 , wherein when executed by the at least one processor of the image display apparatus, the instructions cause the image display apparatus to further perform:

in response to receiving the first audio input for starting the speech recognition from the first audio receiver, initializing a standby time, and

in response to no audio input for speech recognition being received via the first audio receiver until the standby time reaches the first reference time after the start of the speech recognition, restoring the volume of the audio signal output through the audio output unit to a volume level set before the speech recognition starts.

17. The non-transitory computer-readable medium according to claim 11 , wherein when executed by the at least one processor of the image display apparatus, the instructions cause the image display apparatus to further perform:

in response to receiving the first audio input for starting the speech recognition from the first audio receiver, decreasing a volume of an audio signal output through an audio output unit to a predetermined level and initializing a standby time,

after displaying the result of the operation according to the second audio input on the display, re-starting the speech recognition for receiving a fourth audio input for speech recognition associated with the result and initializing the standby time,

in response to no audio input for speech recognition being received via the first receiver until the standby time reaches a second reference time after displaying the result on the display, ending the speech recognition and restoring the volume of the audio signal output through the audio output unit to a volume level set before the speech recognition starts.

18. The non-transitory computer-readable medium according to claim 17 , wherein the second reference time is different from the first reference time.

19. The non-transitory computer-readable medium according to claim 17 , wherein the second reference time is longer than the first reference time.

20. The non-transitory computer-readable medium according to claim 11 , wherein when executed by the at least one processor of the image display apparatus, the instructions cause the image display apparatus to further perform:

controlling a background sound canceller of the image display apparatus to cancel background sound other than the user voice from the first audio receiver.

21. A method of controlling an image display apparatus, the method comprising:

in response to receiving a first audio input for starting a speech recognition from a first audio receiver of the image display apparatus, starting the speech recognition,

in response to a second audio input being received via the first audio receiver after the speech recognition is started, performing an operation according to the second audio input, and controlling a display of the image display apparatus to display a result of the operation, and

in response to a button of a remote controller being pressed while the image display apparatus waits to receive a user voice via the first audio receiver after the speech recognition is started by the first audio input to the first audio receiver of the image display apparatus, ending the speech recognition started by the first audio input,

wherein the remote controller is configured to control the image display apparatus, and the button of the remote controller is configured to control the image display apparatus to start the speech recognition,

wherein the remote controller includes:

a second audio receiver to receive a third audio input wherein the third audio input is received as a user voice via the second audio receiver of the remote controller, and

a communicator to communicate with the image display apparatus and to transmit a control signal to the image display apparatus.

22. The method according to claim 21 , further comprising:

in response to receiving the first audio input for starting the speech recognition from the first audio receiver, initializing a standby time, and

in response to no audio input for speech recognition being received via the first audio receiver until the standby time reaches a first reference time after the start of the speech recognition, ending the speech recognition started by the first audio input.

23. The method according to claim 22 , further comprising:

in response to the second audio input for speech recognition being received via the first audio receiver before the standby time reaches the first reference time after the start of the speech recognition, initializing the standby time.

24. The method according to claim 21 , further comprising:

in response to receiving the first audio input for starting the speech recognition, decreasing a volume of an audio signal output through an audio output unit of the image display apparatus to a predetermined level.

25. The method according claim 24 , wherein the predetermined level is not zero level.

26. The method according to claim 24 , further comprising:

in response to receiving the first audio input for starting the speech recognition from the first audio receiver, initializing a standby time, and

in response to no audio input for speech recognition being received via the first audio receiver until the standby time reaches the first reference time after the start of the speech recognition, restoring the volume of the audio signal output through the audio output unit to a volume level set before the speech recognition starts.

27. The method according to claim 21 , further comprising:

in response to receiving the first audio input for starting the speech recognition from the first audio receiver, decreasing a volume of an audio signal output through an audio output unit to a predetermined level and initializing a standby time,

after displaying the result of the operation according to the second audio input on the display, re-starting the speech recognition for receiving a fourth audio input for speech recognition associated with the result and initializing the standby time,

in response to no audio input for speech recognition being received via the first receiver until the standby time reaches a second reference time after displaying the result on the display, ending the speech recognition and restoring the volume of the audio signal output through the audio output unit to a volume level set before the speech recognition starts.

28. The method according to claim 27 , wherein the second reference time is different from the first reference time.

29. The method according to claim 27 , wherein the second reference time is longer than the first reference time.

30. The method according to claim 21 , further comprising:

controlling a background sound canceller of the image display apparatus to cancel background sound other than the user voice from the first audio receiver.

Priority Claims (2)
KR 10-2012-0002659 · Jan 9, 2012 · national
KR 10-2012-0143590 · Dec 11, 2012 · national
Continuity (7)
Continuation 17167588 · Feb 4, 2021
Continuation 16569849 · Sep 13, 2019
Continuation 15722416 · Oct 2, 2017
Continuation 15351500 · Nov 15, 2016
Continuation 14678556 · Apr 3, 2015
Continuation 13737683 · Jan 9, 2013
Related Publication 20230245653A1 · Aug 3, 2023