Outer Rim Archives
Archives · 2024 · 20240062022

Application (pre-grant publication)

MACHINE TRANSLATION SYSTEM FOR ENTERTAINMENT AND MEDIA

Number
20240062022
Published
2024-02-22
Filed
2023-10-31
Assignee
DISNEY ENTERPRISES, INC.
Inventors
Doggett; Erika
CPC
G06F3/167; G06F40/205; G06F40/40; G06F40/51; G06F40/55; G06F40/58; G10L15/1822; G10L15/22; G10L15/26; G10L21/00; G10L25/57; H04N21/251; H04N21/440236; H04N21/466; H04N21/4856; H04N21/4884; H04N21/8106
Verdict
Set aside generic machine translation, localization plumbing
Source
Google Patents · FreePatentsOnline

Abstract

Techniques for generating translated audio output based on media content are disclosed. Text is accessed corresponding to media content. One or more untranslated mouth shape indicia are determined based on the text. The text is parsed into one or more text chunks when one or more dubbing parameters are met. The parsed text is translated from a first spoken language to a second spoken language. One or more translated mouth shape indicia are determined. The one or more translated mouth shape indicia and the one or more untranslated mouth shape indicia are compared based on a predetermined tolerance threshold. A translated audio output is generated based on the translated text.

Background

BACKGROUND 1. Field

This disclosure generally relates to the field of language translation. 2. General Background

Conventional machine translation systems typically allow for computerized translation from text and/or audio in a first language (e.g., English) into text and/or audio of a second language (e.g., Spanish). For example, some machine translation systems allow for a word-for-word translation from a first language into a second language. Yet, such systems typically focus only on pure linguistic translation. SUMMARY

In one aspect, a computer program product comprises a non-transitory computer readable storage device having a computer readable program stored thereon. The computer readable program when executed on a computer causes the computer to receive, with a processor, audio corresponding to media content. Further, the computer is caused to convert, with the processor, the audio to text. In addition, the computer is caused to concatenate, with the processor, the text with one or more time codes. The computer is also caused to parse, with the processor, the concatenated text into one or more text chunks according to one or more subtitle parameters. Further, the computer is caused to automatically translate, with the processor, the parsed text from a first spoken language to a second spoken language. Moreover, the computer is caused to determine, with the processor, if the language translation complies with the one or more subtitle parameters. Add

Claims

1. At least one non-transitory computer-readable medium carrying instructions that, when executed by a processor, cause the processor to: receive audio corresponding to media content; convert the audio to text; concatenate the text with one or more time codes and one or more untranslated mouth shape indicia; parse the concatenated text into one or more text chunks when one or more dubbing parameters are met; automatically translate the one or more text chunks from a first spoken language to a second spoken language; automatically generate one or more translated mouth shape indicia; determine whether the one or more translated mouth shape indicia match the one or more untranslated mouth shape indicia within a predetermined tolerance threshold; and generate a translated audio output based on the one or more translated mouth shape indicia matching the one or more untranslated mouth shape indicia within the predetermined tolerance threshold. || 8. A computing system comprising: at least one processor; and at least one non-transitory memory carrying instructions that, when executed by the at least one processor, cause the computing system to: access text corresponding to media content; determine one or more untranslated mouth shape indicia based on the text; parse the text into one or more text chunks when one or more dubbing parameters are met; translate the parsed text from a first spoken language to a second spoken language; determine one or more translated mouth shape indicia; compare the one or more translated mouth shape indicia and the one or more untranslated mouth shape indicia based on a predetermined tolerance threshold; and generate a translated audio output based on the translated text. || 16. A computer-implemented method comprising: accessing text corresponding to media content; determining one or more untranslated mouth shape indicia based on the text; parsing the text into one or more text chunks when one or more dubbing parameters are met; translating the parsed text from a first spoken language to a second spoken language; determining one or more translated mouth shape indicia; comparing the one or more translated mouth shape indicia and the one or more untranslated mouth shape indicia based on a predetermined tolerance threshold; and generating a translated audio output based on the translated text.