IP Library Patent Application 11326339
Patent Application
App. No. 11/326,339

Imaging device and image output device

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
11/326,339
Abstract

The device of the present invention synthesizes voices or the like with the image data when an image is taken, and can thereby obtain impressive images or prints with high added values. Furthermore, the present invention can also extract a voice of a specific speaker through a voice print decision and convert it to text, and can thereby improve the accuracy of text conversion.

Claims (33)

1 . An imaging device, comprising:

an image pickup device which takes an image of a speaker;

a voice input device which inputs a voice of the speaker;

a voice print registration device which registers a voice print of the speaker;

a voice extraction device which filters the voice input by the voice input device and extracts the voice corresponding to the voice print registered in the voice print registration device;

a text data generation device which converts the extracted voice to text data; and

a recording device which records the image taken by the image pickup device associated with the text data.

2 . The imaging device according to claim 1 , wherein the voice print registration device registers voice prints of a plurality of speakers associated with speaker identification information which identifies the speakers, and

when voices of the plurality of speakers are input, the text data generation device makes the text data distinguishable for each of the speakers.

3 . The imaging device according to claim 1 , further comprising an image/text synthesis device which synthesizes the image with text image data which is the text data converted to an image.

4 . The imaging device according to claim 2 , further comprising an image/text synthesis device which synthesizes the image with text image data which is the text data converted to an image.

5 . The imaging device according to claim 4 , wherein the image/text synthesis device changes at least one of a character font of the text image data, font size, color, background color, character decoration or column setting for each of the speakers.

6 . The imaging device according to claim 5 , further comprising an extracted voice specification device which selects the speaker identification information and specifies a speaker whose voice is to be extracted by the voice extraction device.

7 . The imaging device according to claim 6 , further comprising a speaker direction calculation device which calculates a direction in which the speaker who utters the voice is located based on the input voice,

wherein the image/text synthesis device lays out the text image data on the image based on the direction in which the speaker is located.

8 . The imaging device according to claim 7 , wherein the voice input device is made up of a plurality of microphones, and

the speaker direction calculation device calculates the direction in which the speaker is located based on differences in sound levels of voices input from the plurality of microphones.

9 . The imaging device according to claim 1 , further comprising a text editing device which edits the text data.

10 . The imaging device according to claim 8 , further comprising a text editing device which edits the text data.

11 . An image output device, comprising:

a data input device which inputs an image and text data associated with the image;

an image/text synthesis device which changes, when the text data is converted to text in such a way that words uttered by a plurality of speakers are made distinguishable for each of the speakers, at least one of a character font of the text image data, font size, color, background color, character decoration or column setting for each of the speakers, synthesizes the text image data with the image to create a synthesized image; and

an output device which outputs the synthesized image.

12 . The image output device according to claim 11 , further comprising a text editing device which edits the text data.

13 . The image output device according to claim 11 , wherein the output device is a printer which prints the image.

14 . The image output device according to claim 12 , wherein the output device is a printer which prints the image.

15 . An image output device, comprising:

a data input device which inputs an image and text data associated with the image;

an image/text synthesis device which lays out, when the text data includes information on a direction in which the speaker is located when an image is taken, the text image data on the image based on the direction in which the speaker is located to create a synthesized image; and

an output device which outputs the synthesized image.

16 . The image output device according to claim 15 , further comprising a text editing device which edits the text data.

17 . The image output device according to claim 15 , wherein the output device is a printer which prints the image.

18 . The image output device according to claim 16 , wherein the output device is a printer which prints the image.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 15, 2007
From: FUJIFILM HOLDINGS CORPORATION (FORMERLY FUJI PHOTO FILM CO., LTD.)
To: FUJIFILM CORPORATION
Reel/Frame 018904/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 6, 2006
From: MIYAZAKI, TAKAO
To: FUJI PHOTO FILM CO., LTD.
Reel/Frame 017446/0442 →