Systems and methods for automated image capture assistance and dual camera mode
Systems and methods are provided for enabling improved image capture at a computing device comprising a plurality of cameras. First and second capture streams, from respective first and second cameras of a computing device, are received at the computing device, wherein the first and second cameras face in different directions. A region of the first capture stream to include as an overlay over a portion of the second capture stream is identified. It is determined that a combined frame comprising a frame from the second capture stream with an overlay from the region of the first capture stream meets threshold criterion based on image component analysis, and, in response to the determining, an image based on the combined frame is stored in a non-transitory memory.
1 . A method comprising:
receiving a first capture stream from a first camera of a computing device;
receiving a second capture stream from a second camera of the computing device, wherein the first camera and the second camera face in different directions;
identifying a region of the first capture stream to include as an overlay over a portion of the second capture stream;
mirroring the region of the first capture stream to obtain consistency of lighting between the region of the first capture stream and the second capture stream;
generating the overlay that includes the mirrored region of the first capture stream;
generating a combined frame comprising a frame from the second capture stream with the generated overlay from the region of the first capture stream;
identifying, via image component analysis, a plurality of image components in the combined frame;
comparing the identified components to at least one predetermined image rule;
determining, based on the comparing, that the combined frame meets a threshold criterion; and
in response to the determining, storing an image based on the combined frame in a non-transitory memory.
2 . The method of claim 1 , wherein the method further comprises generating a combined stream comprising the generated overlay from the region of the first capture stream, at a first location, over the portion of the second capture stream.
3 . The method of claim 1 , wherein the at least one predetermined image rule is a golden spiral rule or a rule-of-thirds rule.
4 . The method of claim 1 , wherein the method further comprises:
generating, for output, the combined frame; and
generating, for output, an indicator associated with storing the image.
5 . The method of claim 1 , wherein identifying the region of the first capture stream further comprises:
identifying a person in the first capture stream; and
generating, via chroma key compositing, the person as the overlay.
6 . The method of claim 1 , further comprising:
generating, for output, a user interface element for moving the generated overlay from the region of the first capture stream in the combined frame;
receiving user input associated with the user interface element; and
moving, based on the received input, the generated overlay from the region of the first capture stream from a first location to a second location in the combined frame.
7 . The method of claim 6 , further comprising:
identifying a suggested location for relocating the generated overlay from the region of the first capture stream in the combined frame; and
generating, for output, an indicator associated with the suggested location.
8 . The method of claim 1 , further comprising:
generating, for output, a user interface element for resizing the generated overlay from the region of the first capture stream in the combined frame;
receiving, at the computing device, user input associated with the user interface element; and
resizing, based on the received input, the generated overlay from the region of the first capture stream from a first size to a second size in the combined frame.
9 . The method of claim 1 , wherein:
identifying the region of the first capture stream further comprises identifying a plurality of sub-regions, each sub-region associated with a face of a person;
the method further comprises receiving, at the computing device, input associated with selecting one or more of the sub-regions; and
the overlay comprises the faces associated with the selected sub-regions.
10 . The method of claim 1 , wherein:
determining that the combined frame meets the threshold criterion further comprises performing the determining via a trained machine learning model; and
the non-transitory memory is a memory of one or more of a computing device and a cloud server.
11 . The method of claim 1 , wherein mirroring the region of the first capture stream comprises aligning a first direction of a source of light associated with the overlay with a second direction of a source of light associated with the second capture stream.
12 . A system comprising:
input circuitry configured to:
receive a first capture stream from a first camera of a computing device; and
receive a second capture stream from a second camera of the computing device, wherein the first camera and the second camera face in different directions; and
control circuitry configured to:
identify a region of the first capture stream to include as an overlay over a portion of the second capture stream;
mirror the region of the first capture stream to obtain consistency of lighting between the region of the first capture stream and the second capture stream;
generate the overlay that includes the mirrored region of the first capture stream;
generate a combined frame comprising a frame from the second capture stream with the generated overlay from the region of the first capture stream;
identify, via image component analysis, a plurality of image components in the combined frame;
compare the identified components to at least one predetermined image rule;
determine, based on the comparing, that the combined frame meets a threshold criterion; and
in response to the determining, store an image based on the combined frame in a non-transitory memory.
13 . The system of claim 12 , wherein the control circuitry is configured to generate a combined stream comprising the generated overlay from the region of the first capture stream, at a first location, over the portion of the second capture stream.
14 . The system of claim 12 , wherein the at least one predetermined image rule is a golden spiral rule or a rule-of-thirds rule.
15 . The system of claim 12 , wherein the control circuitry is further configured to:
generate, for output, the combined frame; and
generate, for output, an indicator associated with storing the image.
16 . The system of claim 12 , wherein the control circuitry configured to identify the region of the first capture stream is further configured to:
identify a person in the first capture stream; and
generate, via chroma key compositing, the person as the overlay.
17 . The system of claim 12 , wherein the control circuitry is further configured to:
generate, for output, a user interface element for moving the generated overlay from the region of the first capture stream in the combined frame;
receive user input associated with the user interface element; and
move, based on the received input, the generated overlay from the region of the first capture stream from a first location to a second location in the combined frame.
18 . The system of claim 17 , wherein the control circuitry is further configured to:
identify a suggested location for relocating the generated overlay from the region of the first capture stream in the combined frame; and
generate, for output, an indicator associated with the suggested location.
19 . The system of claim 12 , wherein the control circuitry is further configured to:
generate, for output, a user interface element for resizing the generated overlay from the region of the first capture stream in the combined frame;
receive, at the computing device, user input associated with the user interface element; and
resize, based on the received input, the generated overlay from the region of the first capture stream from a first size to a second size in the combined frame.
20 . The system of claim 12 , wherein:
the control circuitry configured to identify the region of the first capture stream is further configured to identify a plurality of sub-regions, each sub-region associated with a face of a person;
the control circuitry is further configured to receive, at the computing device, input associated with selecting one or more of the sub-regions; and
the overlay comprises the faces associated with the selected sub-regions.
21 . The system of claim 12 , wherein:
the control circuitry configured to determine that the combined frame meets the threshold criterion is further configured to perform the determining via a trained machine learning model; and
the non-transitory memory is a memory of one or more of a computing device or a cloud server.
22 . The system of claim 12 , wherein mirroring the region of the first capture stream comprises aligning a first direction of a source of light associated with the overlay with a second direction of a source of light associated with the second capture stream.