IP Library › Granted Patent US 11,989,232
Granted Patent B2
US 11,989,232 · App. 17/090,939 · Granted May 21, 2024

Generating realistic representations of locations by emulating audio for images based on contextual information

Inventors: Saraswathi Sailaja Perumalla (Visakhapatnam, IN); Shanthan Chamala (Malvern, PA); Venkata Vara Prasad Karri (Visakhapatnam, IN); Sairam Telukuntla (Westford, MA); Sarbajit K. Rakshit (Kolkata, IN)
Assignee: International Business Machines Corporation
G06F16/687G06F16/638G06T19/006H04R23/008
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,989,232
App. No.
17/090,939
Granted
May 21, 2024
Kind
B2
Abstract

Embodiments of the present invention provide methods, computer program products, and systems. Embodiments of the present invention can dynamically generate audio for one or more images associated with a location based on contextual information that satisfies a request. Embodiments of the present invention can embed the generated audio into the one or more images. Embodiments of the present invention can then display the one or more images with the embedded audio on a user device.

Claims (100)

1. A computer-implemented method comprising:

receiving a request to emulate audio for a location having unknown acoustic properties:

in response to receiving the request to emulate audio for a location having unknown acoustic properties, dynamically generating audio for one or more images associated with the location, based on contextual information that satisfies the request;

embedding the generated audio into the one or more images;

emulating audio in a generated user interface that provides a realistic representation of the location using the embedded audio to simulate interactions of generated noises with materials associated with the one more images at different points in time and in varying weather conditions in response to user interactions with the one or more images;

implementing a feedback mechanism that solicits feedback indicating perceived accuracy for the one or more images associated with the emulated audio that simulates interactions of generated noises with materials associated with the one or more images and interactions of users;

receiving solicited feedback indicating a perceived accuracy for the one or more images associated with the emulated audio;

in response to receiving the solicited feedback indicating the perceived accuracy for the one or more images associated with the emulated audio, altering a depiction of the one or more images to increase an accuracy of depictions of the one or more images and corresponding emulated audio;

considering health parameters of a user by altering the location to be within a threshold score for tolerance prior to the user entering the location, based on the health parameters of the user;

generating one or more recommendations that alter the location to be within a threshold score of audio level; and

alerting the user of alternate locations that meet the threshold score for tolerance.

2. The computer-implemented method of claim 1 , wherein dynamically generating audio for one or more images associated with a location based on contextual information that satisfies a request comprises:

prioritizing contextual information associated with the location;

generating one or more images that match the contextual information; and

generating audio associated with the one or more generated images that match the contextual information.

3. The computer-implemented method of claim 2 , further comprising:

simulating different events at the location by altering at least one object of a plurality of identified objects based on contextual information,

wherein an event simulates interactions of users at the location, and

wherein altering at least one object comprises:

generating one or more images to represent the event; and

embedding audio representative of audio emitted by the generated one or more images representing the event.

4. The computer-implemented method of claim 3 , further comprising:

indexing the plurality of identified objects based on acoustic properties of each identified object of the plurality of identified objects.

5. The computer-implemented method of claim 3 , further comprising:

generating one or more graphical icons to be overlaid on the one or more generated images that represents at least one object of the plurality of objects;

overlaying the at least one or more generated graphical icons over a generated image of the one or more generated images displayed on the user device; and

in response to selecting at least one generated graphical icon of the generated one or more graphical icons, playing audio associated with a respective object of the plurality of objects.

6. The computer-implemented method of claim 1 , further comprising:

generating a score that indicates a noise level associated with an object depicted in respective images of the one or more images; and

in response to the generated score meeting or exceeding a threshold score for the noise level, identifying one or more replacement objects that are functionally equivalent;

generating a graphic representation of at least one replacement object of the one or more replacement objects that are functionally equivalent;

embedding audio associated with the generated graphic representation; and

displaying the graphic representation as an interactive object to a user.

7. The computer-implemented method of claim 1 , further comprising:

generating recommendations to place one or more objects to optimize audio coverage based on number of users present and current location of each of the users that are present.

8. The computer-implemented method of claim 1 , further comprising:

generating a score that indicates a material of the materials associated with the one or more images has a noise level greater than a threshold score for the noise level based on user preference; and

recommending an alternative material to replace the material or otherwise alter the material to reduce noise levels.

9. A computer program product comprising:

one or more computer readable storage media and program instructions stored on the one or more computer readable storage media, the program instructions comprising:

program instructions to receive a request to emulate audio for a location having unknown acoustic properties:

program instructions to, in response to receiving the request to emulate audio for a location having unknown acoustic properties, dynamically generate audio for one or more images associated with the location, based on contextual information that satisfies the request;

program instructions to embed the generated audio into the one or more images;

program instructions to emulate audio in a generated user interface that provides a realistic representation of the location using the embedded audio to simulate interactions of generated noises with materials associated with the one more images at different points in time and in varying weather conditions in response to user interactions with the one or more images;

program instructions to implement a feedback mechanism that solicits feedback indicating perceived accuracy for the one or more images associated with the emulated audio that simulates interactions of generated noises with materials associated with the one or more images and interactions of users;

program instructions to receive solicited feedback indicating a perceived accuracy for the one or more images associated with the emulated audio;

program instructions to, in response to receiving the solicited feedback indicating the perceived accuracy for the one or more images associated with the emulated audio, alter a depiction of the one or more images to increase an accuracy of depictions of the one or more images and corresponding emulated audio;

program instructions to consider health parameters of a user by altering the location to be within a threshold score for tolerance prior to the user entering the location, based on the health parameters of the user;

program instructions to generate one or more recommendations that alter the location to be within a threshold score of audio level; and

program instructions to alert the user of alternate locations that meet the threshold score for tolerance.

10. The computer program product of claim 9 , wherein the program instructions to dynamically generating audio for one or more images associated with a location based on contextual information that satisfies a request comprise:

program instructions to prioritize contextual information associated with the location;

program instructions to generate one or more images that match the contextual information; and

program instructions to generate audio associated with the one or more generated images that match the contextual information.

11. The computer program product of claim 10 , wherein the program instructions stored on the one or more computer readable storage media further comprise:

program instructions to simulate different events at the location by altering at least one object of a plurality of identified objects based on contextual information, wherein an event simulates interactions of users at the location, and

wherein the program instructions to alter at least one object comprise:

program instructions to generate one or more images to represent the event; and

program instructions to embed audio representative of audio emitted by the generated one or more images representing the event.

12. The computer program product of claim 11 , wherein the program instructions stored on the one or more computer readable storage media further comprise:

program instructions to index the plurality of identified objects based on acoustic properties of each identified object of the plurality of identified objects.

13. The computer program product of claim 11 , wherein the program instructions stored on the one or more computer readable storage media further comprise:

program instructions to generate one or more graphical icons to be overlaid on the one or more generated images that represents at least one object of the plurality of objects;

program instructions to overlay the at least one or more generated graphical icons over a generated image of the one or more generated images displayed on the user device; and

program instructions to, in response to selecting at least one generated graphical icon of the generated one or more graphical icons, play audio associated with a respective object of the plurality of objects.

14. The computer program product of claim 9 , wherein the program instructions stored on the one or more computer readable storage media further comprise:

program instructions to generate a score that indicates a noise level associated with an object depicted in respective images of the one or more images; and

program instructions to, in response to the generated score meeting or exceeding a threshold score for the noise level, identify one or more replacement objects that are functionally equivalent;

program instructions to generate a graphic representation of at least one replacement object of the one or more replacement objects that are functionally equivalent;

program instructions to embed audio associated with the generated graphic representation; and

program instructions to display the graphic representation as an interactive object to a user.

15. A computer system comprising:

one or more computer processors;

one or more computer readable storage media; and

program instructions stored on the one or more computer readable storage media for execution by at least one of the one or more computer processors, the program instructions comprising:

program instructions to receive a request to emulate audio for a location having unknown acoustic properties;

program instructions to, in response to receiving the request to emulate audio for a location having unknown acoustic properties, dynamically generate audio for one or more images associated with the location, based on contextual information that satisfies the request;

program instructions to embed the generated audio into the one or more images;

program instructions to emulate audio in a generated user interface that provides a realistic representation of the location using the embedded audio to simulate interactions of generated noises with materials associated with the one more images at different points in time and in varying weather conditions in response to user interactions with the one or more images;

program instructions to implement a feedback mechanism that solicits feedback indicating perceived accuracy for the one or more images associated with the emulated audio that simulates interactions of generated noises with materials associated with the one or more images and interactions of users;

program instructions to receive solicited feedback indicating a perceived accuracy for the one or more images associated with the emulated audio;

program instructions to, in response to receiving the solicited feedback indicating a perceived accuracy for the one or more images associated with the emulated audio, alter a depiction of the one or more images to increase an accuracy of depictions of the one or more images and corresponding emulated audio;

program instructions to consider health parameters of a user by altering the location to be within a threshold score for tolerance prior to the user entering the location, based on the health parameters of the user;

program instructions to generate one or more recommendations that alter the location to be within a threshold score of audio level; and

program instructions to alert the user of alternate locations that meet the threshold score for tolerance.

16. The computer system of claim 15 , wherein the program instructions to dynamically generating audio for one or more images associated with a location based on contextual information that satisfies a request comprise:

program instructions to prioritize contextual information associated with the location;

program instructions to generate one or more images that match the contextual information; and

program instructions to generate audio associated with the one or more generated images that match the contextual information.

17. The computer system of claim 16 , wherein the program instructions stored on the one or more computer readable storage media further comprise:

program instructions to simulate different events at the location by altering at least one object of a plurality of identified objects based on contextual information, wherein an event simulates interactions of users at the location, and

wherein the program instructions to alter at least one object comprise:

program instructions to generate one or more images to represent the event; and

program instructions to embed audio representative of audio emitted by the generated one or more images representing the event.

18. The computer system of claim 17 , wherein the program instructions stored on the one or more computer readable storage media further comprise:

program instructions to index the plurality of identified objects based on acoustic properties of each identified object of the plurality of identified objects.

19. The computer system of claim 17 , wherein the program instructions stored on the one or more computer readable storage media further comprise:

program instructions to generate one or more graphical icons to be overlaid on the one or more generated images that represents at least one object of the plurality of objects;

program instructions to overlay the at least one or more generated graphical icons over a generated image of the one or more generated images displayed on the user device; and

program instructions to, in response to selecting at least one generated graphical icon of the generated one or more graphical icons, play audio associated with a respective object of the plurality of objects.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 6, 2020
From: PERUMALLA, SARASWATHI SAILAJA; CHAMALA, SHANTHAN; KARRI, VENKATA VARA PRASAD; TELUKUNTLA, SAIRAM; RAKSHIT, SARBAJIT K
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 054292/0667 →
Continuity (1)
Related Publication 20220147563A1 · May 12, 2022