IP Library Patent Application 16265303
Patent Application
App. No. 16/265,303

METHOD AND APPARATUS FOR INFORMATION INTERACTION

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
16/265,303
Abstract

A method and an apparatus for information interaction are provided. An embodiment of the method includes: obtaining to-be-processed information, the to-be-processed information comprising textual information and an image; extracting a feature word from the textual information of the to-be-processed information, and searching for descriptive information of the image within the to-be-processed information based on the feature word, the feature word being used to characterize a search request for the image, and the descriptive information being used to characterize a textual description of the image; and constructing response information to the to-be-processed information from the descriptive information. The embodiment constructs response information from the descriptive information, thereby enabling the information interaction with the to-be-processed information and improving the efficiency of information interaction.

Claims (44)

1 . A method for information interaction, the method comprising:

obtaining to-be-processed information, the to-be-processed information comprising textual information and an image;

extracting a feature word from the textual information of the to-be-processed information, and searching for descriptive information of the image within the to-be-processed information based on the feature word, the feature word being used to characterize a search request for the image, and the descriptive information being used to characterize a textual description of the image; and

constructing response information to the to-be-processed information from the descriptive information.

2 . The method according to claim 1 , wherein the extracting a feature word from the textual information of the to-be-processed information comprises:

performing a semantic recognition on the textual information to obtain semantic information corresponding to the textual information; and

extracting the feature word from the semantic information.

3 . The method according to claim 1 , wherein the searching for descriptive information of the image within the to-be-processed information based on the feature word comprises:

importing the image into an image search model to obtain a to-be-matched image set corresponding to the image, wherein the to-be-matched image set comprises at least one to-be-matched image, and the image search model is configured to characterize a first corresponding relationship between the image and the to-be-matched image;

importing the to-be-matched image into a semantic tagging model to obtain a semantic tag set corresponding to the to-be-matched image set, wherein the semantic tagging model is configured to characterize a second corresponding relationship between the to-be-matched image and a semantic tag, and the semantic tag is used to provide a textual description of the to-be-matched image; and

selecting a to-be-recognized semantic tag from the semantic tag set, and using interpretive information of a noun in the to-be-recognized semantic tag, the noun corresponding to the image, as the descriptive information.

4 . The method according to claim 3 , wherein the selecting a to-be-recognized semantic tag from the semantic tag set comprises:

counting numbers of identical semantic tags within the semantic tag set, and using the semantic tag having a maximum number as the to-be-recognized semantic tag.

5 . The method according to claim 4 , the method further comprising correcting the descriptive information, wherein the correcting the descriptive information comprises:

receiving feedback information corresponding to the response information, wherein the feedback information is used to evaluate an accuracy of the response information;

performing a semantic recognition on the feedback information to obtain the accuracy;

choosing a secondary to-be-recognized tag from the semantic tags in the semantic tag set excluding the to-be-recognized semantic tag, in response to determining that the accuracy is below a preset threshold;

using the interpretive information of the noun in the secondary to-be-recognized tag, the noun corresponding to the image, as secondary descriptive information; and

constructing the response information to the to-be-processed information from the secondary descriptive information.

6 . An apparatus for information interaction, the apparatus comprising:

at least one processor; and

a memory storing instructions, the instructions when executed by the at least one processor, cause the at least one processor to perform operations, the operations comprising:

obtaining to-be-processed information, the to-be-processed information comprising textual information and an image;

extracting a feature word from the textual information of the to-be-processed information, and searching for descriptive information of the image within the to-be-processed information based on the feature word, the feature word being used to characterize a search request for the image, and the descriptive information being used to characterize a textual description of the image; and

constructing response information to the to-be-processed information from the descriptive information.

7 . The apparatus of according to claim 6 , wherein the extracting a feature word from the textual information of the to-be-processed information comprises:

performing a semantic recognition on the textual information to obtain semantic information corresponding to the textual information; and

extracting the feature word from the semantic information.

8 . The apparatus according to claim 6 , wherein the searching for descriptive information of the image within the to-be-processed information based on the feature word comprises:

importing the image into an image search model to obtain a to-be-matched image set corresponding to the image, wherein the to-be-matched image set comprises at least one to-be-matched image, and the image search model is configured to characterize a first corresponding relationship between the image and the to-be-matched image;

importing the to-be-matched image into a semantic tagging model to obtain a semantic tag set corresponding to the to-be-matched image set, wherein the semantic tagging model is configured to characterize a second corresponding relationship between the to-be-matched image and a semantic tag, and the semantic tag is used to provide a textual description of the to-be-matched image; and

selecting a to-be-recognized semantic tag from the semantic tag set, and using interpretive information of a noun in the to-be-recognized semantic tag, the noun corresponding to the image, as the descriptive information.

9 . The apparatus according to claim 8 , wherein the selecting a to-be-recognized semantic tag from the semantic tag set comprises:

counting numbers of identical semantic tags within the semantic tag set, and using the semantic tag having a maximum number as the to-be-recognized semantic tag.

10 . The apparatus according to claim 9 , the operations further comprising correcting the descriptive information, wherein the correcting the descriptive information comprises:

receiving feedback information corresponding to the response information, wherein the feedback information is used to evaluate an accuracy of the response information;

performing a semantic recognition on the feedback information to obtain the accuracy;

choosing a secondary to-be-recognized tag from the semantic tags in the semantic tag set excluding the to-be-recognized semantic tag, in response to determining that the accuracy is below a preset threshold;

using the interpretive information of the noun in the secondary to-be-recognized tag, the noun corresponding to the image, as secondary descriptive information; and

constructing the response information to the to-be-processed information from the secondary descriptive information.

11 . A non-transitory computer-readable storage medium storing a computer program, the computer program when executed by one or more processors, causes the one or more processors to perform operations, the operations comprising:

obtaining to-be-processed information, the to-be-processed information comprising textual information and an image;

extracting a feature word from the textual information of the to-be-processed information, and searching for descriptive information of the image within the to-be-processed information based on the feature word, the feature word being used to characterize a search request for the image, and the descriptive information being used to characterize a textual description of the image; and

constructing response information to the to-be-processed information from the descriptive information.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 30, 2021
From: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.; SHANGHAI XIAODU TECHNOLOGY CO. LTD.
Reel/Frame 056811/0772 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 4, 2019
From: TIAN, XIAOLI; FANG, GAOLIN; GU, XIAOGUANG; MI, XUE; SUN, KE; DING, XINZHE; SUN, RUIYING
To: BAIDU ONLINE NETWORK TECHNOLOGY (BEIJING) CO., LTD.
Reel/Frame 048233/0479 →