IP Library Granted Patent US 9,626,360
Granted Patent B2
US 9,626,360 · App. 13/933,815 · Granted Apr 18, 2017

Analyzing web site for translation

Inventors: Enrique Travieso (Davie, FL); Adam Rubenstein (Boca Raton, FL); William Fleming (Boca Raton, FL)
Assignee: MOTIONPOINT CORPORATION
G06F17/289G06F17/30861H04L67/02Y10S707/99942Y10S707/99943Y10S707/99945Y10S707/99948Y10S707/99953Y10S707/99955
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,626,360
App. No.
13/933,815
Granted
Apr 18, 2017
Kind
B2
Abstract

A system, method and computer readable medium for synchronizing web content is disclosed. The method includes retrieving a first web content in a first language from a web site, the first web content corresponding to a second web content wherein the second web content is a translation in a second language of the first web content. The method further includes dividing the first web content into a plurality of translatable components and generating a unique identifier for each of the plurality of translatable components. The method further includes matching each of the plurality of translatable components to a plurality of translated components of the second web content using the unique identifier of each of the plurality of translatable components. If a translatable component is not matched to a translated component, the method further includes designating the translatable component for translation into the second language.

Claims (28)

1. A machine implemented method for providing statistics characterizing translation work in synchronizing content in different languages, comprising the steps of:

activating a spider agent to crawl a website for content in a first language via a publicly accessible network path for synchronizing a translated version of the content in the first language previously generated in a second language;

parsing the crawled content in the first language into a plurality of translatable components;

accessing a database that stores translated components previously generated in the second language;

identifying at least some of the plurality of translatable components in the first language that do not have a corresponding translated component in the second language in the database;

generating statistics based on the identified translatable components to estimate the work load involved in language translation of the identified translatable components from the first language to the second language; and

providing information including the generated statistics to characterize a service related to synchronizing the content in the first and second languages.

2. The method according to claim 1 , wherein the language translation includes human translating the identified translatable components.

3. The method according to claim 1 , further comprising the step of adding the identified translatable components to a translation list for translation into a second language.

4. The method according to claim 1 , further comprising the step of generating an identifier for each of the plurality of translatable components such that each of the plurality of translatable components is accessible via a corresponding identifier.

5. The method according to claim 4 , wherein the identifier for a text segment is generated using at least one of a hash code, a checksum, and a mathematical algorithm based on one or more text segments.

6. The method according to claim 1 , wherein the statistics includes at least one of a file count, a page count, a text segment count, a unique text segment count, a word count, and a unique word count.

7. The method of claim 1 , wherein the step of generating comprises:

computing the statistics based on information associated with any of the identified translatable components that does not have a corresponding translated component in the second language.

8. A machine readable non-transitory medium having information stored thereon for providing statistics characterizing translation work in synchronizing content in different languages, wherein the information, when read, causes the machine to perform the following:

activating a spider agent to crawl a website for content in a first language via a publicly accessible network path for synchronizing a translated version of the content in the first language previously generated in a second language;

parsing the crawled content in the first language into a plurality of translatable components;

accessing a database that stores translated components previously generated in the second language;

identifying at least some of the plurality of translatable components in the first language that do not have a corresponding translated component in the second language in the database;

generating statistics based on the identified translatable components to estimate the work load involved in language translation of the identified translatable components from the first language to the second language; and

providing information including the statistics to characterize a service related to synchronizing the content in the first and second languages.

9. The medium according to claim 8 , wherein the language translation includes human translating the identified translatable components.

10. The medium according to claim 8 , wherein the information, when read, further causes the machine to perform the following: adding the identified translatable components to a translation list for translation into a second language.

11. The medium according to claim 8 , wherein the information, when read, further causes the machine to perform the following: generating an identifier for each of the plurality of translatable components such that each of the plurality of translatable components is accessible via a corresponding identifier.

12. The medium according to claim 11 , wherein the identifier for a text segment is generated using at least one of a hash code, a checksum, and a mathematical algorithm based on one or more text segments.

13. The medium according to claim 8 , wherein the statistics includes at least one of a file count, a page count, a text segment count, a unique text segment count, a word count, and a unique word count.

14. The medium according to claim 8 , wherein the step of generating comprises:

computing the statistics based on information associated with any of the identified translatable components that does not have a corresponding translated component in the second language.

Assignments (4)
SECURITY INTEREST Recorded Mar 31, 2021
From: MOTIONPOINT CORPORATION
To: CADENCE BANK, N.A.
Reel/Frame 055784/0020 →
SECURITY INTEREST Recorded Mar 31, 2021
From: MOTIONPOINT CORPORATION
To: MARANON CAPITAL, L.P., AS ADMINISTRATIVE AGENT
Reel/Frame 055787/0130 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 17, 2021
From: TRAVIESO, ENRIQUE; RUBENSTEIN, ADAM; FLEMING, WILLIAM
To: MOTIONPOINT CORPORATION
Reel/Frame 055623/0989 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 2, 2017
From: TREVIESO, ENRIQUE; RUBENSTEIN, ADAM; FLEMING, WILLIAM
To: MOTIONPOINT CORPORATION
Reel/Frame 041443/0388 →
Continuity (4)
Continuation 12609834 · Oct 30, 2009
Continuation 10784334 · Feb 23, 2004
Provisional Application 60449571 · Feb 21, 2003
Related Publication 20140058719A1 · Feb 27, 2014