IP Library Granted Patent US 8,838,616
Granted Patent B2
US 8,838,616 · App. 12/543,263 · Granted Sep 16, 2014

Server device for creating list of general words to be excluded from search result

Inventor: Norikazu Matsumura (Tokyo, JP)
Assignee: NEC Biglobe, Ltd.
G06F17/30684G06F17/30731
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,838,616
App. No.
12/543,263
Granted
Sep 16, 2014
Kind
B2
Abstract

A server device of the present invention includes a control unit collecting texts stored in a storage unit in response to an instruction from the outside or when a predetermined time is reached, extracting words from the collected texts, determining, as a general word, a word which appears at a frequency higher than a first predefined value for a first predetermined period, and which appears at a frequency that varies within a second predefined value range for every second predetermined period that is shorter than the first predetermined period, and creating a general word list which enumerates the general words.

Claims (66)

1. A server device comprising:

a control unit collecting texts stored in a storage unit in response to a general word extracting request signal input by a user or when a predetermined time is reached, extracting words from said collected texts, determining, as a general word, a word which appears at a frequency higher than a first predefined value for a first predetermined period, and which appears at a frequency varying within a second predefined value range for every second predetermined period shorter than said first predetermined period, and creating a general word list which enumerates a plurality of general words including said general word,

wherein said control unit, when a keyword for a search is entered, collects texts including said keyword from texts stored in said storage unit, extracts nouns from collected first texts, determines a noun which partially matches said keyword as a first word, extracts second texts including said first word from among said first texts, extracts a word which belongs to at least one word from among a noun, verb, and adjective from said second texts, counts a number of times said extracted word is used, determines words which are ranked at a predetermined position or higher with respect to the number of times said words are used, as second words which are pertinent word to said first word, lowers the rank of a second word which matches any one from among said plurality of general words in said general word list, and outputs said second words together with said first word.

2. A server device comprising:

a control unit collecting texts stored in a storage unit in response to a general word extracting request signal input by a user or when a predetermined time is reached, extracting a plurality of words from said collected texts, calculating a score for each of said plurality of words based on an appearance frequency for a first predetermined period and an appearance frequency for every second predetermined period shorter than said first predetermined period, and creating a general word list which includes said plurality of words and said scores,

wherein said control unit, when a keyword for a search is entered, collects texts including said keyword from texts stored in said storage unit, extracts nouns from collected first texts, determines a noun which partially matches said keyword as a first word, extracts second texts including said first word from among said first texts, extracts a word which belongs to at least one word from among a noun, verb, and adjective from said second texts, counts number of times said extracted word is used, determines words which are ranked at a predetermined position or higher with respect to the number of times said words are used, as second words which are pertinent word to said first word, lowers the rank of a second word which matches any one from among said plurality of words in said general word list, and outputs said second words together with said first word.

3. The server device according to claim 1 , wherein said every second predetermined period is daily, weekly, or monthly.

4. The server device according to claim 2 , wherein said every second predetermined period is daily, weekly, or monthly.

5. The server device according to claim 1 , wherein said appearance frequency for said first predetermined period is of one type of category comprising a number of articles in which said word appears for the period, or a proportion of the number of articles in which said word appears, or a ranking of number of articles in which said word appears, and said appearance frequency for said second predetermined period is the number of articles in which said word appears for the period, or a proportion of the number of articles in which said word appears, or a ranking of the number of articles in which said word appears in correspondence to said category of said appearance frequency for said first predetermined period.

6. The server device according to claim 2 , wherein said appearance frequency for said first predetermined period is of one type of category comprising a number of articles in which said word appears for the period, or a proportion of the number of articles in which said word appears, or a ranking of number of articles in which said word appears, and said appearance frequency for said second predetermined period is the number of articles in which said word appears for the period, or a proportion of the number of articles in which said word appears, or a ranking of the number of articles in which said word appears in correspondence to said category of said appearance frequency for said first predetermined period.

7. A server device comprising:

a control unit collecting texts stored in a storage unit in response to a general word extracting request signal input by a user or when a predetermined time is reached, extracting words from said collected texts, determining, as a general word, a word which appears at a frequency higher than a first predefined value for a first predetermined period, and which appears at a frequency varying within a second predefined value range for every second predetermined period shorter than said first predetermined period, and creating a general word list which enumerates a plurality of general words including said general word,

wherein said control unit, when a keyword for a search is entered, collects texts including said keyword from texts stored in said storage unit, extracts nouns from collected first texts, determines a noun which partially matches said keyword as a first word, extracts second texts including said first word from among said first texts, extracts a word which belongs to at least one word from among a noun, verb, and adjective from said second texts, counts a number of times said extracted word is used, determines words which are ranked at a predetermined position or higher with respect to the number of times said words are used, as second words which are pertinent word to said first word, deletes a second word which matches any one from among said plurality of general words in said general word list, and outputs remained second words together with said first word.

8. A server device comprising:

a control unit collecting texts stored in a storage unit in response to a general word extracting request signal input by a user or when a predetermined time is reached, extracting a plurality of words from said collected texts, calculating a score for each of said plurality of words based on an appearance frequency for a first predetermined period and an appearance frequency for every second predetermined period shorter than said first predetermined period, and creating a general word list which includes said plurality of words and said scores,

wherein said control unit, when a keyword for a search is entered, collects texts including said keyword from texts stored in said storage unit, extracts nouns from collected first texts, determines a noun which partially matches said keyword as a first word, extracts second texts including said first word from among said first texts, extracts a word which belongs to at least one word from among a noun, verb, and adjective from said second texts, counts a number of times said extracted word is used, determines words which are ranked at a predetermined position or higher with respect to the number of times said words are used, as second words which are pertinent word to said first word, deletes a second word which matches any one from among said plurality of words in said general word list, and outputs remained second words together with said first word.

9. An information processing method comprising:

collecting texts stored in a storage unit in response to a general word extracting request signal input by a user or when a predetermined time is reached, and extracting words from said collected texts;

determining, as a general word, a word which appears at a frequency higher than a first predefined value for a first predetermined period, and which appears at a frequency varying within a second predefined value range for every second predetermined period shorter than said first predetermined period;

creating a general word list which enumerates a plurality of general words including said general word;

collecting texts including a keyword from texts stored in said storage unit in response to said keyword entered for a search;

extracting nouns from collected first texts, determining a noun which partially matches said keyword as a first word;

extracting second texts including said first word from among said first texts;

extracting a word which belongs to at least one word from among a noun, verb, and adjective from said second texts;

counting a number of times said word extracted from said second texts is used;

determining words extracted from said second texts, as second words which are pertinent word to said first word, if the words are ranked at a predetermined position or higher with respect to the number of times the words are used; and

lowering the rank of a second word which matches any one from among said plurality of general words in said general word list, and outputting said second words together with said first word.

10. An information processing method comprising:

collecting texts stored in a storage unit in response to a general word extracting request signal input by a user or when a predetermined time is reached, and extracting a plurality of words from said collected texts;

calculating a score for each of said plurality of words based on an appearance frequency for a first predetermined period and an appearance frequency for every second predetermined period shorter than said first predetermined period;

creating a general word list which includes said plurality of words and said scores;

collecting texts including a keyword from texts stored in said storage unit in response to said keyword entered for a search;

extracting nouns from collected first texts, determining a noun which partially matches said keyword as a first word;

extracting second texts including said first word from among said first texts;

extracting a word which belongs to at least one word from among a noun, verb, and adjective from said second texts;

counting a number of times said word extracted from said second texts is used;

determining words extracted from said second texts, as second words which are pertinent word to said first word, if the words are ranked at a predetermined position or higher with respect to the number of times the words are used; and

lowering the rank of a second word which matches any one from among said plurality of words in said general word list, and outputting said second words together with said first word.

11. The information processing method according to claim 9 , wherein said every second predetermined period is daily, weekly, or monthly.

12. The information processing method according to claim 10 , wherein said every second predetermined period is daily, weekly, or monthly.

13. The information processing method according to claim 9 , wherein said appearance frequency for the first predetermined period is of one type of category comprising a number of articles in which said word appears for the period, or a proportion of the number of articles in which said word appears, or a ranking of the number of articles in which said word appears, and said appearance frequency for said second predetermined period is the number of articles in which said word appears for the period, or a proportion of the number of articles in which said word appears, or a ranking of the number of articles in which said word appears in correspondence to said category of said appearance frequency for said first predetermined period.

14. The information processing method according to claim 10 , wherein said appearance frequency for the first predetermined period is of one type of category comprising a number of articles in which said word appears for the period, or a proportion of the number of articles in which said word appears, or a ranking of the number of articles in which said word appears, and said appearance frequency for said second predetermined period is the number of articles in which said word appears for the period, or a proportion of the number of articles in which said word appears, or a ranking of the number of articles in which said word appears in correspondence to said category of said appearance frequency for said first predetermined period.

15. An information processing method comprising:

collecting texts stored in a storage unit in response to a general word extracting request signal input by a user or when a predetermined time is reached, extracting words from said collected texts;

determining, as a general word, a word which appears at a frequency higher than a first predefined value for a first predetermined period, and which appears at a frequency varying within a second predefined value range for every second predetermined period shorter than said first predetermined period;

creating a general word list which enumerates a plurality of general words including said general word;

collecting texts including a keyword from texts stored in said storage unit in response to said keyword entered for a search;

extracting nouns from collected first texts;

determining a noun which partially matches said keyword as a first word;

extracting second texts including said first word from among said first texts;

extracting a word which belongs to at least one word from among a noun, verb, and adjective from said second texts;

counting a number of times said word extracted from said second texts is used;

determining words extracted from said second texts, as second words which are pertinent word to said first word, if the words are ranked at a predetermined position or higher with respect to the number of times the words are used; and

deleting a second word which matches any one from among said plurality of general words in said general word list, and outputting remained second words together with said first word.

16. An information processing method comprising:

collecting texts stored in a storage unit in response to a general word extracting request signal input by a user or when a predetermined time is reached, and extracting a plurality of words from said collected texts;

calculating a score for each of said plurality of words based on an appearance frequency for a first predetermined period and an appearance frequency for every second predetermined period shorter than said first predetermined period;

creating a general word list which includes said plurality of words and said scores;

collecting texts including a keyword from texts stored in said storage unit in response to said keyword entered for a search;

extracting nouns from collected first texts;

determining a noun which partially matches said keyword as a first word;

extracting second texts including said first word from among said first texts;

extracting a word which belongs to at least one word from among a noun, verb, and adjective from said second texts;

counting a number of times said word extracted from said second texts is used;

determining words extracted from said second texts, as second words which are pertinent word to said first word, if the words are ranked at a predetermined position or higher with respect to the number of times the words are used; and

deleting a second word which matches any one from among said plurality of words in said general word list, and outputting remained second words together with said first word.

Assignments (3)
CHANGE OF ADDRESS Recorded Aug 4, 2015
From: BIGLOBE INC.
To: BIGLOBE INC.
Reel/Frame 036263/0822 →
CHANGE OF NAME Recorded Oct 27, 2014
From: NEC BIGLOBE, LTD.
To: BIGLOBE INC.
Reel/Frame 034056/0365 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 18, 2009
From: MATSUMURA, NORIKAZU
To: NEC BIGLOBE, LTD.
Reel/Frame 023114/0063 →
Priority Claims (1)
JP 2008-216465 · Aug 26, 2008 · national
Continuity (1)
Related Publication 20100057724A1 · Mar 4, 2010