Outer Rim Archives
Archives · 2017 · 20170039191

Application (pre-grant publication)

LINGUISTIC ANALYSIS AND CORRECTION

Number
20170039191
Published
2017-02-09
Filed
2015-08-06
Assignee
Disney Enterprises, Inc.
Inventors
Sakashita; Tadashi George et al.
CPC
G06F40/51; G06F40/129
Verdict
Set aside linguistic analysis and correction, localization/NLP tooling
Source
Google Patents · FreePatentsOnline

Abstract

Methods, computer program products, and systems for correcting a glyph in a translated text are described. In one embodiment, the method includes identifying a first form of a first glyph in a translation text having a plurality of contextual properties and analyzing, by the processor, the first form of the first glyph with reference to one or more glyph form tables comprising a plurality of forms of the first glyph based, at least in part, on the plurality of contextual properties.

Background

BRIEF DESCRIPTION OF THE DRAWINGS

FIG. 1 is a functional block diagram of a linguistic evaluation system.

FIG. 2 is a flowchart illustrating a method of correcting location based glyph form errors.

FIG. 3 is a flowchart illustrating a method of correcting contextual based glyph form errors.

FIG. 4 is an example of a location based glyph form table.

FIG. 5A is an example of computer program instructions for correcting location based glyph form errors.

FIG. 5B is an example of computer program instructions for correcting glyph form errors.

FIG. 6 is a functional block diagram of exemplary components of the linguistic evaluation system of FIG. 1.OVERVIEW

Embodiments described herein are directed to the identification and correction of errors in translations to contextual based languages employing glyphs whose forms and/or meanings depends on certaincontextual properties, such as the locations and/or contexts in which they are used. As used herein, “contextual based language” means any language that includes one or more glyphs having different forms depending on the context in which it is used. Examples of contextual based languages include Arabic, Hebrew, Chinese, and Japanese, among others. Contextual based languages can be contrasted with non-contextual based languages, such as English, in which the form of a letter is independent of the context in which it is used (e.g., theletter “a” always appears the same regardless of the

Claims

1. A method for analyzing contextual language content comprising: identifying, by a processor, a first form of a first glyph having a plurality of contextual properties in the contextual language content; and analyzing, by the processor, the first form of the first glyph with reference to one or more glyph form tables comprising a plurality of forms of the first glyph based, at least in part, on the plurality of contextual properties. 9. A computer program product for correcting a symbol in a translated text comprising: a non-transitory, machine readable storage device having computer program instructions stored thereon for execution by a processor, the program instructions comprising: program instructions to identify a symbol having a plurality of contextual properties in a translation text; program instructions to compare a first appearance of the symbol to one or more tables describing a plurality of appearances of the symbol based, at least in part, on theplurality of contextual properties; program instructions to determine a second appearanceof the symbol based, at least in part, on the one or more tables and the contextual properties; and program instructions to replace the first appearance of the symbol with the second appearance of the symbol in the translated text. 16. A system for correcting a character in a translated text comprising: one or more processors; and a non-transitory, machine readable storage device having computer program instructions stored thereon for execution by a processor, the program instructions comprising: program instructions to identify a first style of a character having a plurality of contextual properties in a translation text; program instructions to compare the character to one or more data structures stored in the storage device and describing a plurality of styles of the character based, at least inpart, on the plurality of contextual properties; program instructions to determine a second style of the character based, at least in part, on the one or more data structures and the contextual properties; and program instructions to replace the first style of the character with the second style of the character in the translated text.