Outer Rim Archives
Archives · 2022 · 20220358701

Application (pre-grant publication)

Emotion-Based Sign Language Enhancement of Content

Number
20220358701
Published
2022-11-10
Filed
2021-10-20
Assignee
Disney Enterprises, Inc.
Inventors
Brandon; Marc, Arana; Mark
CPC
G06F40/169; H04N21/4884; G06V20/40; G06F40/58; G06F40/30; G06T13/00; G06F40/216; H04N21/4316; G10L21/055; G06V20/41; G06F40/20; H04N21/44008; G06T11/00; H04N21/43079; H04N21/43074; G10L25/57; H04N21/440236; G10L25/63; G10L15/22; H04N21/4394; G09B21/009; G06F3/1423; G06V40/174
Verdict
Set aside sign-language content-enhancement feature, accessibility/business
Source
Google Patents · FreePatentsOnline

Abstract

A content enhancement system includes a computing platform having processing hardware and a system memory storing software code. The processing hardware is configured to execute the software code to receive audio-video (A/V) content, to execute at least one of a visual analysis or an audio analysis of the A/V content, and to determine, based on executing the at least one of the visual analysis or the audio analysis, an emotional aspect of the A/V content. The processing hardware is further configured to execute the software code to generate, using the emotional aspect of the A/V content, a sign language translation of the A/V content, the sign language translation including one or more of a gesture, a posture, or a facial expression conveying the emotional aspect.

Background

BACKGROUND

Members of the deaf and hearing impaired communities often rely on any of a number of signed languages for communication via hand signals. Although effective in translating the plain meaning of a communication, hand signals alone typically do not fully capture the emphasis or emotional intensity motivating that communication. Accordingly, skilled human sign language translators tend to employ multiple physical modes when communicating information. Those modes may include gestures other than hand signals, postures, and facial expressions, as well as the speed and force with which such expressive movements are executed.

For a human sign language translator, identification of the appropriate emotional intensity and emphasis to include in a signing performance may be largely intuitive, based on cognitive skills honed unconsciously as the understanding of spoken language is learned and refined through childhood and beyond. However, the exclusive reliance on human sign language translation can be expensive, and in some use cases may be inconvenient or even impracticable. Consequently, there is a need in the art for an automated solution for providing emotion-based sign language enhancement of content.

Claims

1. A content enhancement system comprising: a computing platform including a processing hardware and a system memory storing a software code: the processing hardware configured to execute the software code to: receive audio-video (A/V) content; execute at least one of a visual analysis or an audio analysis of the A/V content; determine, based on executing the at least one of the visual analysis or the audio analysis, an emotional aspect of the A/V content; and generate, using the emotional aspect of the A/V content, a sign language translation of the A/V content, the sign language translation eluding one or more of a gesture, a posture, or a facial expression conveying the emotional aspect. || 2. The content enhancement system of claim further comprising: a display; wherein the processing hardware is further configured to execute the software code to: render the A/V content on the display; and render a performance of the sign language translation on the display concurrently with rendering the A/V content corresponding to the sign language translation. || 11. A method for use by a content enhancement system including a computing platform having a processing hardware and a system memory storing a software code, the method comprising: receiving, by the software code executed by the processing hardware, audio-video content; executing, by the soft are code executed by the processing hardware, at least one of a visual analysis or an audio analysis of the A/V content; determining, by the software code executed by the processing hardware based on executing the at least one of the visual analysis or the audio analysis, an emotional aspect of the A/V content; and generating, by the software code executed by the processing hardware, using the emotional aspect of the A/V content, a sign language translation of the A/V content, the sign language translation including one or more of a gesture, a posture, or a facial expression conveying the emotional aspect.