Outer Rim Archives
Archives · 2022 · 20220345234

Application (pre-grant publication)

SYSTEM FOR DELIVERABLES VERSIONING IN AUDIO MASTERING

Number
20220345234
Published
2022-10-27
Filed
2021-04-21
Assignee
Lucasfilm Entertainment Company Ltd. LLC
Inventors
Morris; Stephen, Levine; Scott, Tsingos; Nicolas
CPC
H04S3/002; H04H60/04; G11B27/031; G06V20/46; G10L25/57; G06F3/165
Verdict
Set aside audio-production pipeline tool, business
Source
Google Patents · FreePatentsOnline

Abstract

Some implementations of the disclosure relate to using a model trained on mixing console data of sound mixes to automate the process of sound mix creation. In one implementation, a non-transitory computer-readable medium has executable instructions stored thereon that, when executed by a processor, causes the processor to perform operations comprising: obtaining a first version of a sound mix; extracting first audio features from the first version of the sound mix obtaining mixing metadata; automatically calculating with a trained model, using at least the mixing metadata and the first audio features, mixing console features; and deriving a second version of the sound mix using at least the mixing console features calculated by the trained model.

Background

BRIEF SUMMARY OF THE DISCLOSURE

Implementations of the disclosure describe systems and methods that leverage machine learning to automate the process of creating various versions of sound mixes.

In one embodiment, a non-transitory computer-readable medium has executable instructions stored thereon that, when executed by a processor, causes the processor to perform operations comprising: obtaining a first version of a sound mix; extracting first audio features from the first version of the sound mix obtaining mixing metadata; automatically calculating with a trained model, using at least the mixing metadata and the first audio features, mixing console features; and deriving a second version of the sound mix using at least the mixing console features calculated by the trained model.

In some implementations, deriving the second version of the sound mix, comprises: inputting the mixing console features derived by the trained model into a mixing console for playback; and recording an output of the playback.

In some implementations, deriving the second version of the sound mix, comprises: displaying to a user, in a human readable format, one or more of the mixing console features derived by the trained model. In some implementations, deriving the second version of the sound mix, further comprises: receiving data corresponding to one or more modifications input by the user modifying one or more of the displayed mixing console features derived by the train

Claims

1. A non-transitory computer-readable medium having executable instructions stored thereon that, when executed by a processor, causes the processor to perform operations comprising: obtaining a first version of a sound mix; extracting first audio features from the first version of the sound mix obtaining mixing metadata; automatically calculating with a trained model, using at least the mixing metadata and the first audio features, mixing console features; and deriving a second version of the sound mix using at least the mixing console features calculated by the trained model. || 13. A non-transitory computer-readable medium having executable instructions stored thereon that, when executed by a processor, causes the processor to perform operations comprising: obtaining a first version of a sound mix; extracting first audio features from the first version of the sound mix extracting video features from video corresponding to the first version of the sound mix; obtaining mixing metadata; and automatically calculating with a trained model, using at least the mixing metadata, the first audio features, and the video features: second audio features corresponding to a second version of the sound mix; or pulse-code modulation (PCM) audio or coded audio corresponding to a second version of the sound mix. || 16. A sound mixing system, comprising: one or more processors; and one or more non-transitory computer-readable mediums having executable instructions stored thereon that, when executed by the one or more processors, cause the one or more processors to perform operations comprising: obtaining a first version of a sound mix; extracting first audio features from the first version of the sound mix obtaining mixing metadata; automatically calculating with a trained model, using at least the mixing metadata and the first audio features, mixing console features; and deriving a second version of the sound mix using at least the mixing console features calculated by the trained model.