Outer Rim Archives
Archives · 2019 · 10255714

Granted patent

System and method of gaze predictive rendering of a focal area of an animation

Number
10255714
Published
2019-04-09
Filed
2016-08-24
Assignee
Disney Enterprises, Inc.
Inventors
Mitchell; Kenneth J., Andrews; Sheldon, Cosker; Darren, Swafford; Nicholas T.
CPC
G06F3/013; G06T13/20; G06T15/20; G06F3/012; G06F3/011
Verdict
Medium Notable software
Source
Google Patents · FreePatentsOnline

The keeper's note

Gaze predictive rendering focal area of animation (foveated rendering).

Abstract

Individual images for individual frames of an animation may be rendered to include individual focal areas. A focal area may include one or more of a foveal region corresponding to a gaze direction of a user, an area surrounding the foveal region, and/or other components. The foveal region may comprise a region along the user's line of sight that permits high visual acuity with respect to a periphery of the line of sight. A focal area within an image may be rendered based on parameter values of rendering parameters that are different from parameter values for an area outside the focal area.

Background

FIELD OF THE DISCLOSURE(1) This disclosure relates to a system and method of gaze-predictive rendering of a focal area of an animation.BACKGROUND(2) When rendering digital images in animations, it is often assumed that the human visual system is perfect, despite limitations arising from a variety of different complexities and phenomena. That is, current methods of real-time rendering of a digital animation may operate on an assumption that a single rendered frame image will be fully visually appreciated at any single point in time. However, peripheral vision may be significantly worse than foveal vision in many ways, and these differences may not be explained solely by a loss of acuity. However, acuity sensitivity still forms a significant portion of peripheral detail loss and can be a phenomena to exploit.(3) One method of exploitation, termed “foveated rendering” or “foveated imaging,” implements a high-resolution render of a particular region of individual frame images. A users gaze may be tracked so that the high-resolution render is positioned on the images to correspond with a user's foveal region. An area surrounding the high-resolution region is then rendered at relatively lower resolution. However, users may experience visual anomaly when prompted about the fact. Other techniques have implemented a foveated rendering method with spatial and temporal property variation. With such techniques, at a certain level-of-detail (LOD), users may experience the foveated renders

Claims

1. A system configured for gaze-predictive rendering of a focal area of an animation presented on a display, wherein the animation includes a sequence of frames, the sequence of frames including a first frame and a plurality of subsequent frames, the system comprising: one or more physical processors configured by machine-readable instructions to: obtain state information describing a state of a virtual space, the state at an individual point in time defining one or more virtual objects within the virtual space and their positions; determine a field of view of the virtual space, the frames of the animation being images of the virtual space within the field of view, such that the first frame is an image of the virtual space within the field of view at a point in time that corresponds to the first frame; apply the state information for a subsequent frame in the plurality of subsequent frames for the animation to a machine learning model to generate a prediction, the prediction relating to one or more expected saccades and one or more gaze directions of a user currently viewing the presented animation, wherein the machine learning model is trained based on statistical targets of eye fixation corresponding to the user's foveal region, and wherein the statistical targets of eye fixation are precomputed on a database of previous eye tracked viewing sessions of the animation; determine, prior to rendering the subsequent frame, the focal area of the subsequent frame within the field of view based on the prediction, such that the focal area includes a foveal region corresponding to the one or more expected saccades and the one or more gaze directions and includes one or more regions outside of the foveal region, wherein the foveal region is a region along the user's line of sight that permits high visual acuity with respect to a periphery of the line of sight; and render, from the state information, one or more images for the subsequent frame of the animation, the one or more images depicting the virtual space within the field of view determined at individual points in time, wherein an area outside of the focal area of the subsequent frame is rendered at a lower resolution than that of the focal area to reduce latency. 10. A method of gaze-predictive rendering of a focal area of an animation presented on a display, the animation including a sequence of frames, the sequence of frames including a first frame and a plurality of subsequent frames, the method being implemented in a computer system comprising one or more physical processors and storage media storing machine-readable instructions, the method comprising: obtaining state information describing state of a virtual space, the state at an individual point in time defining one or more virtual objects within the virtual space and their positions; determining a field of view of the virtual space, the frames of the animation being images of the virtual space within the field of view, such that the first frame is an image of the virtual space within the field of view at a point in time that corresponds to the first frame; applying the state information for a subsequent frame in the plurality of subsequent frames for the animation to a machine learning model to generate a prediction, the prediction relating to one or more expected saccades and one or more gaze directions of a user currently viewing the presented animation, wherein the machine learning model is trained based on statistical targets of eye fixation corresponding to the user's foveal region, and wherein the statistical targets of eye fixation are precomputed on a database of previous eye tracked viewing sessions of the animation; determining, prior to rendering the subsequent frame, the focal area of the subsequent frame within the field of view based on the prediction, such that the focal area includes a foveal region corresponding to the one or more expected saccades and the one or more gaze directions and includes one or more regions outside of the foveal region, wherein the foveal region is a region along the user's line of sight that permits high visual acuity with respect to a periphery of the line of sight; and rendering from the state information, one or more images for the subsequent frame of the animation, the one or more images depicting the virtual space within the field of view determined at individual points in time, wherein an area outside of the focal area of the subsequent frame is rendered at a lower resolution than that of the focal area to reduce latency.