Shared point of view
Sound sources can be spatially rendered in a setting and shown through a display. In response to satisfaction of a threshold criterion that is satisfied based on relative distance between the sound sources and a position of a listener, the rendering of the sound sources can be adjusted to maintain spatial integrity of the sound sources. The adjustment can be performed to prevent one of the sound sources from arriving at the listener earlier than another of the sound sources.
1 . A method, comprising:
spatially rendering a first sound source to be perceived by a listener at a first position relative to a virtual listener position in a setting that is shown through a display to the listener;
spatially rendering a second sound source to be perceived at a second position relative to the virtual listener position in the setting; and
in response to satisfaction of a threshold criterion that includes at least one of: a) a distance between the virtual listener position and the first position, b) a distance between the virtual listener position and the second position, and c) a difference of the distance between the virtual listener position and the second position and the distance between the virtual listener position and the first position,
adjusting the spatial rendering of the first sound source so that the first sound source is to be perceived by the listener at a third position relative to the virtual listener position that is closer to the virtual listener position than the first position in the setting.
2 . The method of claim 1 , wherein the first position is shared with a computer generated avatar in the setting, and wherein the first sound source contains voice captured at a physical location.
3 . The method of claim 2 , wherein the second position is shared with a virtual display showing a stream of images captured at the physical location, and wherein the second sound source contains ambience captured at the physical location.
4 . The method of claim 3 , wherein the second sound source contains residual of the voice or the first sound source contains residual of the ambience.
5 . The method of claim 1 , wherein the first position is shared with a virtual display showing a stream of images captured at a physical location, and wherein the first sound source contains ambience captured at the physical location.
6 . The method of claim 1 , wherein the second position is shared with a computer generated avatar in the setting, and wherein the second sound source comprises voice captured at a physical location.
7 . The method of claim 6 , wherein the second sound source contains residual of ambience or the first sound source contains residual of the voice.
8 . The method of claim 1 , wherein the threshold criterion further includes at least one of: d) an amount of residual of the first sound source contained in the second sound source, e) an amount of residual of the second sound source contained in the first sound source, and f) a loudness of the first sound source or the second sound source.
9 . A method, comprising:
spatially rendering a first sound source to be perceived by a listener at a first position relative to a virtual listener position in a setting that is shown through a display to the listener;
spatially rendering a second sound source to be perceived at a second position relative to the virtual listener position in the setting; and
in response to satisfaction of a threshold criterion that includes at least one of: a) a distance between the virtual listener position and the first position, b) a distance between the virtual listener position and the second position, and c) a difference of the distance between the virtual listener position and the second position and the distance between the virtual listener position and the first position,
applying spatial filters to the second sound source so that the second sound source is to be perceived by the listener after the first sound source is to be perceived by the listener regardless of if the virtual listener position becomes closer to the second sound source.
10 . The method of claim 9 , wherein the first position is shared with a computer generated avatar in the setting, and wherein the first sound source contains voice captured at a physical location.
11 . The method of claim 10 , wherein the second position is shared with a virtual display showing a stream of images captured at the physical location, and wherein the second sound source contains ambience captured at the physical location.
12 . The method of claim 11 , wherein the second sound source contains residual of the voice or the first sound source contains residual of the ambience.
13 . The method of claim 9 , wherein the first position is shared with a virtual display showing a stream of images captured at a physical location, and wherein the first sound source contains ambience captured at the physical location.
14 . The method of claim 9 , wherein the second position is shared with a computer generated avatar in the setting, and wherein the second sound source comprises voice captured at a physical location.
15 . The method of claim 14 , wherein the second sound source contains residual of ambience or the first sound source contains residual of the voice.
16 . The method of claim 9 , wherein the threshold criterion further includes at least one of: d) an amount of residual of the first sound source contained in the second sound source, e) an amount of residual of the second sound source contained in the first sound source, and f) a loudness of the first sound source or the second sound source.
17 . A method, comprising:
spatially rendering a first sound source to be perceived by a listener at a first position relative to a virtual listener position in a setting that is shown through a display to the listener;
spatially rendering a second sound source to be perceived at a second position relative to the virtual listener position in the setting; and
in response to satisfaction of a threshold criterion that includes at least one of: a) a distance between the virtual listener position and the first position, b) a distance between the virtual listener position and the second position, and c) a difference of the distance between the virtual listener position and the second position and the distance between the virtual listener position and the first position,
moving the first sound source or the second sound source in the setting such that the threshold criterion is no longer satisfied, or preventing a change to the virtual listener position if the change would result in the threshold criterion being satisfied.
18 . The method of claim 17 , wherein the first position is shared with a computer generated avatar in the setting, and wherein the first sound source contains voice captured at a physical location.
19 . The method of claim 18 , wherein the second position is shared with a virtual display showing a stream of images captured at the physical location, and wherein the second sound source contains ambience captured at the physical location.
20 . The method of claim 19 , wherein the second sound source contains residual of the voice or the first sound source contains residual of the ambience.
21 . The method of claim 17 , wherein the first position is shared with a virtual display showing a stream of images captured at a physical location, and wherein the first sound source contains ambience captured at the physical location.
22 . The method of claim 17 , wherein the second position is shared with a computer generated avatar in the setting, and wherein the second sound source comprises voice captured at a physical location.
23 . The method of claim 22 , wherein the second sound source contains residual of ambience or the first sound source contains residual of the voice.
24 . The method of claim 17 , wherein the threshold criterion further includes at least one of: d) an amount of residual of the first sound source contained in the second sound source, e) an amount of residual of the second sound source contained in the first sound source, and f) a loudness of the first sound source or the second sound source.