Surfacing dynamic call transcripts for a secondary display
A method of enabling dynamic surfacing of transcripts of communication exchanges between first and second communicatively connected electronic communication devices. The method includes, during a communication exchange between the first and second electronic communication devices, determining whether a trusted person is detected within a threshold distance in proximity of a first display of the first electronic device. The method includes, in response to detecting a trusted person within the threshold distance, determining whether the trusted person is in an inquisitive mode relative to at least one of a user of the first electronic device and the communication exchange. The method includes, in response to determining that the trusted person is in an inquisitive mode, activating a transcript sharing mode of the first electronic device and starting surfacing of a transcript of an ongoing communication on the first display of the first electronic device.
1 . An electronic device comprising:
an enclosure comprising a first surface and a second surface opposed to the first surface;
a communications subsystem comprising at least one interface by which the electronic device communicatively connects via a wireless connection to a second electronic device to enable a user of the electronic device to engage in a communication exchange with a second user of the second electronic device;
at least one display comprising a first display incorporated into the first surface, and oriented to be outward facing, away from the user of the electronic device while the user is engaged in the communication exchange, the electronic device to selectively generate and present call transcripts on the first display;
at least one image capturing device having a field of view (FOV) in viewing range of the first display;
and
at least one processor communicatively coupled to the communication subsystem, the first display, and the image capturing device, the at least one processor executing program code, and configured to cause the electronic device to:
during the communication exchange, determine whether a trusted person is detected within a threshold distance in proximity of the first display of the electronic device;
in response to detecting a trusted person within the threshold distance, determine whether the trusted person is in an inquisitive mode relative to at least one of a first user of the electronic device and the communication exchange; and
in response to determining that the detected trusted person is within the threshold distance and in an inquisitive mode, activate a transcript sharing mode of the electronic device and start surfacing a transcript of an ongoing communication on the first display of the electronic device.
2 . The electronic device of claim 1 , wherein the at least one processor is further configured to cause the electronic device to:
determine whether the first electronic device is operating in a private environment; and
initiate the presenting of the transcript on the first display only if the first electronic device is operating in a private environment.
3 . The electronic device of claim 2 , wherein to determine whether the electronic device is operating in a private environment, the processor causes the electronic device to:
capture, by the image capturing device, images of persons within the FOV, the FOV including an area within which faces of persons are visible in a viewable range of the first display;
compare the images to stored images of trusted persons; and
initiate a surfacing of the transcript only when every person identified within the viewable range is a trusted person.
4 . The electronic device of claim 3 , wherein each trusted person is pre-designated as such via an assigned trusted contact stored with identifying information comprising at least an associated name and at least one identifying image.
5 . The electronic device of claim 1 , further comprising:
a second display, incorporated into the second surface of the enclosure, and oriented towards the user of the first electronic device while the user is engaged in the communication exchange;
a microphone that enables capture of audible input for the communication exchange; and
an audio output system that enables receipt by the user of audio from the communication exchange.
6 . The electronic device of claim 5 , further comprising:
at least one depth sensor that enables detection of three dimensionality of an object whose image is captured within the FOV of the at least one image capturing device, the depth sensor being operative to check for depth using a hardware-based active depth perception sensing technology;
wherein the at least one processor analyzes data received from the at least one depth sensor to determine whether a preview image detected by the at least one image capturing device is a face of a physical person, wherein the at least one processor is configured to cause the electronic device to differentiate between a still image and an actual physical person facing the first display of the electronic device and to activate transcript sharing mode only in response to the detected object being an actual physical person.
7 . The electronic device of claim 5 , wherein the at least one processor is configured to cause the electronic device to:
present the transcript on the first display; and
concurrently present the transcript on a first section of the second display along with a communication exchange user interface.
8 . The electronic device of claim 1 , further comprising a second display incorporated into the second surface facing the user, wherein the at least one processor is further configured to cause the electronic device to activate the transcript sharing mode by generating and presenting, via the second display, a selectable activation option for receiving user approval prior to the surfacing of the transcript on the first display.
9 . The electronic device of claim 1 , wherein the at least one processor is further configured to cause the electronic device to generate and present the transcript comprising a verbatim textual representation of words spoken by a user of the second electronic device during the communication exchange.
10 . The electronic device of claim 1 , wherein the at least one processor is further configured to cause the electronic device activate a dedicated artificial intelligence (AI) module that generates and presents the transcript comprising a summary, generated by the AI module, of a context associated with words spoken by both a user of the second electronic device and the user of the electronic device during the communication exchange.
11 . A method comprising:
during a communication exchange between a first electronic device and a communicatively connected second electronic device, determining, by a processor of the first electronic device, whether a trusted person is detected within a threshold distance in proximity of a first display of the first electronic device, the first electronic device comprising:
an enclosure comprising a first surface and a second surface opposed to the first surface;
a communications subsystem comprising at least one interface by which the first electronic device communicatively connects via a wireless connection to the second electronic device to enable a user of the first electronic device to engage in a communication exchange with a second user of the second electronic device;
at least one display comprising a first display incorporated into the first surface, and oriented to be outward facing, away from the user of the electronic device while the user is engaged in the communication exchange, the electronic device to selectively generate and present call transcripts on the first display; and
at least one image capturing device having a field of view (FOV) in viewing range of the first display;
in response to detecting a trusted person within the threshold distance, determining whether the trusted person is in an inquisitive mode relative to at least one of a user of the first electronic device and the communication exchange; and
in response to determining that the trusted person is in an inquisitive mode, activating a transcript sharing mode of the first electronic device and surfacing a transcript of an ongoing communication on the first display of the first electronic device.
12 . The method of claim 11 , wherein, in determining whether the electronic device is operating in a private environment, the method further comprises:
capturing, by an image capturing device, images of persons within an FOV that are in a viewable range of the first display;
comparing the images to stored images of trusted persons; and
initiating a surfacing of the transcript only when every person identified within the viewable range is a trusted person.
13 . The method of claim 12 , further comprising:
determining, via at least one depth sensor, a three dimensionality of an object whose image is captured within the FOV of the image capturing device, the depth sensor being operative to check for depth using a hardware-based active depth perception sensing technology; and
analyzing data received from the at least one depth sensor to determine whether a preview image detected by the at least one image capturing device is a face of a physical person, wherein a processor is configured to cause the first electronic device to differentiate between a still image and an actual physical person facing the first display of the first electronic device and to activate transcript sharing mode only in response to the detected object being an actual physical person.
14 . The method of claim 11 , further comprising: activating the transcript sharing mode by generating and presenting, via a second display of the first electronic device facing the user, a selectable activation option for user selection prior to the surfacing of the transcript on the first display.
15 . The method of claim 14 , further comprising:
presenting the transcript on the first display; and
concurrently presenting the transcript on a first section of the second display along with a communication exchange user interface.
16 . The method of claim 11 , further comprising generating and presenting the transcript comprising a verbatim written representation of words spoken by a user of the second electronic device during the communication exchange.
17 . The method of claim 11 , further comprising generating, by a dedicated artificial intelligence (AI) module, the transcript comprising a summary of a context associated with words spoken by both a user of the second electronic device and the user of the first electronic device during the communication exchange.
18 . A computer program product comprising a non-transitory computer readable medium having program instructions that when executed by a processor of a first electronic device that comprises an interface that enables the first electronic device to communicatively connect via a wireless connection to a second electronic device to effect a communication exchange, configure the first electronic device to perform functions comprising:
during a communication exchange between the first electronic device and a communicatively connected second electronic device, determining whether a trusted person is detected within a threshold distance in proximity of a first display of the first electronic device, the first electronic device comprising:
an enclosure comprising a first surface and a second surface opposed to the first surface;
a communications subsystem comprising the interface by which the first electronic device communicatively connects via the wireless connection to the second electronic device to enable a user of the first electronic device to engage in the communication exchange with a second user of the second electronic device;
a first display incorporated into the first surface, and oriented to be outward facing, away from the user of the electronic device while the user is engaged in the communication exchange, the electronic device to selectively generate and present call transcripts on the first display; and
at least one image capturing device having a field of view (FOV) in viewing range of the first display;
in response to detecting a trusted person within the threshold distance, determining whether the trusted person is in an inquisitive mode relative to at least one of a first user of the first electronic device and the communication exchange; and
in response to determining that the trusted person is in an inquisitive mode, activating a transcript sharing mode of the first electronic device and surfacing a transcript of an ongoing communication on the first display of the first electronic device.
19 . The computer program product of claim 18 , further comprising program instructions for, in determining whether the first electronic device is operating in a private environment:
capturing, by an image capturing device, images of persons within an FOV that are in a viewable range of the first display;
comparing the images to stored images of trusted persons; and
initiating a surfacing of the transcript only when every person identified within the viewable range is a trusted person.
20 . The computer program product of claim 19 , further comprising program instructions for:
determining, via at least one depth sensor, a three dimensionality of an object whose image is captured within the FOV of the image capturing device; and
analyzing data received from the at least one depth sensor to determine whether a preview image detected by the at least one image capturing device is a face of a physical person, wherein the at least one processor is configured to cause the first electronic device to differentiate between a still image and an actual physical person facing the first display of the first electronic device and to activate transcript sharing mode only in response to the detected object being an actual physical person.