Outer Rim Archives
Archives · 2022 · 11523188

Granted patent

Systems and methods for intelligent media content segmentation and analysis

Number
11523188
Published
2022-12-06
Filed
2016-06-30
Assignee
Disney Enterprises, Inc.
Inventors
Narayan; Nimesh, Luu; Jack, Pao; Alan, Petrillo; Matthew, Accardo; Anthony M., Lindquist; Alexis, Farre Guiu; Miquel Angel, Ettinger; Katharine S., Volodarsky Bareket; Lena
CPC
G06V20/49; H04N21/47205; G11B27/28; H04N21/8133; H04N21/8456
Verdict
Set aside content segmentation analytics, business
Source
Google Patents · FreePatentsOnline

Abstract

There is provided a system including a non-transitory memory storing an executable code and a hardware processor executing the executable code to receive a media content including a plurality of frames, divide the media content into a plurality of shots, each of the plurality of shots including a plurality of frames of the media content based on a first similarity between the plurality of frames, determine a plurality of sequential shots of the plurality of shots to be part of a first sub-scene of a plurality of sub-scenes of a scene based on a timeline continuity of the plurality of sequential shots, identify each of the plurality of shots of the media content and each of the plurality of sub-scenes with a corresponding beginning time code and a corresponding ending time code.

Background

BACKGROUND (1) Typical video programs, such as television shows and movies, include a number of different video shots and scenes shown in sequence, the content of which may be processed using video content analysis. Conventional video content analysis may be utilized to identify motion in a video, recognize objects and/or shapes in a video, and track an object or a person in a video. SUMMARY (2) The present disclosure is directed to systems and methods for intelligent media content segmentation and analysis, substantially as shown in and/or described in connection with at least one of the figures, as set forth more completely in the claims.

Claims

1. A system comprising: a non-transitory memory storing an executable code; a hardware processor configured to execute the executable code to: receive a media content; divide the media content into a plurality of shots, each of the plurality of shots including a plurality of frames of the media content, wherein the media content is divided into the plurality of shots based on a first similarity between the plurality of frames; determine a plurality of sequential shots of the plurality of shots to be part of a first sub-scene of a plurality of sub-scenes of a scene based on a timeline continuity of the plurality of sequential shots; identify each of the plurality of shots of the media content and each of the plurality of sub-scenes with a corresponding beginning time code and a corresponding ending time code; receive a user input annotating at least one of the identified plurality of shots or the identified plurality of sub-scenes; and store the user input in an annotation database. || 9. A method for use with a system comprising a non-transitory memory and a hardware processor, the method comprising: receiving, using the hardware processor, a media content; dividing, using the hardware processor, the media content into a plurality of shots, each of the plurality of shots including a plurality of frames of the media content, wherein the media content is divided into the plurality of shots based on a first similarity between the plurality of frames; determining, using the hardware processor, a plurality of sequential shots of the plurality of shots to be part of a first sub-scene of a plurality of sub-scenes of a scene based on a timeline continuity of the plurality of sequential shots; and identifying, using the hardware processor, each of the plurality of shots of the media content and each of the plurality of sub-scenes with a corresponding beginning time code and a corresponding ending time code; receiving, using the hardware processor, a user input annotating at least one of the identified plurality of shots or the identified plurality of sub-scenes; and storing, using the hardware processor, the user input in an annotation database.