- Number
- 12626434
- Published
- 2026-05-12
- Filed
- 2024-01-29
- Assignee
- DISNEY ENTERPRISES, INC.
- Inventors
- Iglesias Navarro; Santiago, Pernías Pascual de Pobil; Pablo, Moore; Robert B., Juboor; David N.
- CPC
- G06T11/60; G06N20/00
- Verdict
- Low Notable software
- First reported
- 2026-W29 (2026-07-15)
- Source
- Google Patents · FreePatentsOnline
The keeper's note
Techniques for generating modified images using facial content information are disclosed.
Abstract
Techniques for generating modified images using facial content information are disclosed. First image data comprising first facial content information is received, and a facial content encoder generates a first embedding by extracting the first facial content information from the first image data. Second image data comprising second facial content information and non-facial content information (e.g., style information, pose, facial expression) is received, and a non-facial content encoder generates a second embedding comprising the non-facial content information. A decoder generates a modified image using the first embedding and the second embedding, the modified image comprising the first facial content information of the first image data and the non-facial content information of the second image data.
Background
CROSS REFERENCE TO RELATED APPLICATION (1) This application is related to the Applicant's concurrently filed application titled “Image Style Transfer,” which is incorporated herein by reference in its entirety for all purposes. FIELD (2) Described embodiments relate generally to generating modified images, such as modifying facial information in an image. BACKGROUND (3) Digital images can be modified in various ways to generate modified images. For example, images can be digitally manipulated to add or remove content or to replace a person's likeness with that of a different person. Modified images can also be generated to combine characteristics or content of images. Image manipulations can be applied manually or using various algorithms. Current techniques may have only limited functionality to transfer selected information, e.g., facial information, or combine information from different images in a fast and accurate manner. SUMMARY (4) The following Summary is for illustrative purposes only and does not limit the scope of the technology disclosed in this document. (5) In an embodiment, a computer-implemented method of generating modified images using facial content is disclosed. First image data is received including first facial content information. The first facial content information can include shapes or dimensions of a set of facial features in the first image data. A first embedding (e.g., a facial content embedding) is generated using a facial content encoder, the e
Claims
1. A computing system comprising: at least one processor; and at least one non-transitory memory carrying instructions that, when executed by the at least one processor, cause the computing system to perform operations comprising: receive first image data comprising first facial content information; generate, using a facial content encoder, a facial content embedding comprising the first facial content information, wherein the facial content encoder extracts the first facial content information from the first image data to generate the facial content embedding; receive second image data comprising second facial content information and non-facial content information; generate, using a non-facial content encoder, a non-facial content embedding comprising the non-facial content information; generate, by a decoder, a modified image using the facial content embedding and the non-facial content embedding, wherein the modified image comprises the first facial content information of the first image data and the non-facial content information of the second image data. ||
7. A non-transitory computer-readable medium carrying instructions that, when executed by a computing system, cause the computing system to perform operations comprising: receive first image data comprising first facial content information; generate, using a facial content encoder, a facial content embedding comprising the first facial content information, wherein the facial content encoder extracts the first facial content information from the first image data to generate the facial content embedding; receive second image data comprising second facial content information and non-facial content information; generate, using a non-facial content encoder, a non-facial content embedding comprising the non-facial content information; generate, by a decoder, a modified image using the facial content embedding and the non-facial content embedding, wherein the modified image comprises the first facial content information of the first image data and the non-facial content information of the second image data. ||
14. A computer-implemented method of generating modified images using facial content information, the method comprising: receiving first image data comprising first facial content information; generating, using a facial content encoder, a first embedding comprising the first facial content information, wherein the facial content encoder extracts the first facial content information from the first image data to generate the first embedding; receiving second image data comprising second facial content information and non-facial content information; generating, using a non-facial content encoder, a second embedding comprising the non-facial content information; generating, by a decoder, a modified image using the first embedding and the second embedding, wherein the modified image comprises the first facial content information of the first image data and the non-facial content information of the second image data.