US2017177567A1PendingUtilityA1

Analyzing Web Site for Translation

Assignee: MOTIONPOINT CORPPriority: Feb 21, 2003Filed: Mar 2, 2017Published: Jun 22, 2017
Est. expiryFeb 21, 2023(expired)· nominal 20-yr term from priority
G06F 40/58G06F 40/44G06F 40/205Y10S707/99943G06F 16/95Y10S707/99945Y10S707/99955Y10S707/99953H04L 67/02Y10S707/99948Y10S707/99942G06F 17/2705G06F 17/2247G06F 17/289G06F 17/2818G06F 40/143
65
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

A system, method and computer readable medium for synchronizing web content is disclosed. The method includes retrieving a first web content in a first language from a web site, the first web content corresponding to a second web content wherein the second web content is a translation in a second language of the first web content. The method further includes dividing the first web content into a plurality of translatable components and generating a unique identifier for each of the plurality of translatable components. The method further includes matching each of the plurality of translatable components to a plurality of translated components of the second web content using the unique identifier of each of the plurality of translatable components. If a translatable component is not matched to a translated component, the method further includes designating the translatable component for translation into the second language.

Claims

exact text as granted — not AI-modified
What is claimed is: 
     
         1 . A method, implemented on a machine having at least one processor, storage, and a communication platform, for providing statistics characterizing translation work in synchronizing content in different languages, comprising:
 receiving a request from a user for accessing content hosted on a website in a first language, wherein the user requests to view the content in a second language and at least some of the content in the first language has previously been translated into the second language;   obtaining the content in the first language from the website via a publicly accessible network path based on the request;   parsing the obtained content in the first language into a plurality of translatable components;   accessing a database that stores the content in the second language previously translated as translated components;   identifying at least some of the plurality of translatable components that do not have a corresponding translated component in the database;   generating statistics based on the at least some of the translatable components to estimate the work load involved in translation of the at least some of the translatable components from the first language to the second language; and   providing the statistics to characterize a service related to synchronizing the content in the first and second languages.   
     
     
         2 . The method according to  claim 1 , wherein the translation includes human translating the at least some of the plurality of translatable components. 
     
     
         3 . The method according to  claim 1 , further comprising adding the at least some of the plurality of translatable components to a translation list for translation into the second language. 
     
     
         4 . The method according to  claim 1 , further comprising generating an identifier for each of the plurality of translatable components such that each of the plurality of translatable components is accessible via a corresponding identifier. 
     
     
         5 . The method according to  claim 4 , wherein the identifier for a text segment is generated using at least one of a hash code, a checksum, and a mathematical algorithm based on one or more text segments. 
     
     
         6 . The method according to  claim 1 , wherein the statistics includes at least one of a file count, a page count, a text segment count, a unique text segment count, a word count, and a unique word count. 
     
     
         7 . The method of  claim 1 , wherein the generating comprises:
 computing the statistics based on information associated with any of the at least some of the plurality of translatable components that do not have a corresponding translated component in the second language.   
     
     
         8 . A machine readable non-transitory medium having information stored thereon for providing statistics characterizing translation work in synchronizing content in different languages, wherein the information, when read, causes the machine to perform the following:
 receiving a request from a user for accessing content hosted on a website in a first language, wherein the user requests to view the content in a second language and at least some of the content in the first language has previously been translated into the second language;   obtaining the content in the first language from the website via a publicly accessible network path based on the request;   parsing the obtained content in the first language into a plurality of translatable components;   accessing a database that stores the content in the second language previously translated as translated components;   identifying at least some of the plurality of translatable components that do not have a corresponding translated component in the database;   generating statistics based on the at least some of the translatable components to estimate the work load involved in translation of the at least some of the translatable components from the first language to the second language; and   providing the statistics to characterize a service related to synchronizing the content in the first and second languages.   
     
     
         9 . The medium according to  claim 8 , wherein the translation includes human translating the at least some of the plurality of translatable components. 
     
     
         10 . The medium according to  claim 8 , wherein the information, when read, further causes the machine to perform the following: adding the at least some of the plurality of translatable components to a translation list for translation into the second language. 
     
     
         11 . The medium according to  claim 8 , wherein the information, when read, further causes the machine to perform the following: generating an identifier for each of the plurality of translatable components such that each of the plurality of translatable components is accessible via a corresponding identifier. 
     
     
         12 . The medium according to  claim 11 , wherein the identifier for a text segment is generated using at least one of a hash code, a checksum, and a mathematical algorithm based on one or more text segments. 
     
     
         13 . The medium according to  claim 8 , wherein the statistics includes at least one of a file count, a page count, a text segment count, a unique text segment count, a word count, and a unique word count. 
     
     
         14 . The medium according to  claim 8 , wherein the generating comprises:
 computing the statistics based on information associated with any of the at least some of the plurality of translatable components that do not have a corresponding translated component in the second language.   
     
     
         15 . A system having at least one processor, storage, and a communication platform, for providing statistics characterizing translation work in synchronizing content in different languages, wherein the at least one processor is configured for:
 receiving a request from a user for accessing content hosted on a website in a first language, wherein the user requests to view the content in a second language and at least some of the content in the first language has previously been translated into the second language;   obtaining the content in the first language from the website via a publicly accessible network path based on the request;   parsing the obtained content in the first language into a plurality of translatable components;   accessing a database that stores the content in the second language previously translated as translated components;   identifying at least some of the plurality of translatable components that do not have a corresponding translated component in the database;   generating statistics based on the at least some of the translatable components to estimate the work load involved in translation of the at least some of the translatable components from the first language to the second language; and   providing the statistics to characterize a service related to synchronizing the content in the first and second languages.

Join the waitlist — get patent alerts

Track US2017177567A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.