Application (pre-grant publication)
SYSTEM AND METHOD TO PROVIDE PERSONALIZED AUDIO STREAMING AND RENDERING
- Number
- 20260236215
- Published
- 2026-08-13
- Filed
- 2025-04-17
- Assignee
- Disney Enterprises, Inc.
- Inventors
- Briand; Manuel
- CPC
- G06F3/165; G10L21/0364; H03G3/32; H03G9/005; H03G9/025; H04N21/4394; H04N21/4852
- Verdict
- Set aside streaming, codec, cdn/infra
- In edition
- 2026-W36
- Source
- Google Patents · FreePatentsOnline
The keeper's note
Systems and methods to provide personalized audio streaming and rendering include receiving a cinematic audio track for selected content and a maximum accessible audio track for the content, and adjustably combining the…
Abstract
Systems and methods to provide personalized audio streaming and rendering include receiving a cinematic audio track for selected content and a maximum accessible audio track for the content, and adjustably combining the cinematic audio track and accessible audio track to provide an improved dialogue audio track which allows a user/listener to hear the dialogue over an environmental noise floor, the combining being provided by a cross-fade renderer adjusted manually by the user/listener or automatically adjusted based on a measured noise floor, and optionally providing a personalized equalizer for hearing impairments. It also allows a content provider to create the maximum accessible audio track for a given content separate from the user/listener, which may also use a X-Fade renderer to verify quality of the accessible audio track.
Background
BACKGROUND
Current audio streaming solutions typically rely on a single audio bitstream and decoder per streaming application which limits the ability of the content owners and streaming service providers to personalize the audio experience it provides to the end user.
In particular, existing steaming services provide separate streams for different versions enhanced dialogue, e.g., English dialogue boost high, English dialogue boost medium, and the like. In that case, the user must select which version they want to listen to. Also, to change to another audio version, e.g., because the background environment noise changed, the user must make a request, and a new version is retrieved from the appropriate content server. Additionally, most TV devices apply post-processing to the decoded audio bitstream, such as AI sound enhancements, to reduce noise or boost dialogue which does not preserve the integrity of the original sound mix nor the creative intent of the content owners.
Switching audio streams can be cumbersome, slow and incur digital streaming file buffering/loading delays during the audio delivery from the Content Delivery Network (CDN) to the end user. Also, conventional systems only provide a pre-defined number of dialogue enhanced versions (e.g., 3-5 versions), without any guarantee that a given selected version may be ideal for the noise environment of the listener/user. Such an approach can make it difficult for the user to hear the dialogue of