Outer Rim Archives
Archives · 2018 · 10019992

Granted patent

Speech-controlled actions based on keywords and context thereof

Number
10019992
Published
2018-07-10
Filed
2015-06-29
Assignee
Disney Enterprises, Inc.
Inventors
Lehman; Jill Fain et al.
CPC
G10L15/22
Verdict
Set aside generic keyword/context speech-controlled actions, no device context
Source
Google Patents · FreePatentsOnline

Abstract

A device includes a plurality of components, a memory having a keyword recognition module and a context recognition module, a microphone configured to receive an input speech spoken by a user, an analog-to-digital converter configured to convert the input speech from an analog form to a digital form and generate a digitized speech, and a processor. The processor is configured to detect, using the keyword recognition module, a keyword in the digitized speech, initiate, in response to detecting the keyword by the keyword recognition module, an action to be taken one of the plurality of components, wherein the keyword is associated with the action, determine, using the context recognition module, a context for the keyword, and execute the action if the context determined by the context recognition module indicates that the keyword is a command.

Background

BACKGROUND(1) As speech recognition technology has advanced, voice-activated devices have become more and more popular and have found new applications. Today, an increasing number of mobile phones, in-home devices, and automobile devices include speech or voice recognition capabilities. Although the speech recognition modules incorporated into such devices are trained to recognize specific keywords, they tend to be unreliable. This is because the specific keywords may appear in a spoken sentence and be incorrectly recognized as voice commands by the speech recognition module when not intended by the user. Also, in some cases, the specific keywords intended to be taken as commands may not be recognized by the speech recognition module, because the specific keywords may appear in between other spoken words, and be ignored. Both situations can frustrate the user and cause the user to give up and resort to inputting the commands manually, speak the keywords numerous times or turn off the voice recognition.SUMMARY(2) The present disclosure is directed to speech-controlled actions based on keywords and context thereof, substantially as shown in and/or described in connection with at least one of the figures, as set forth more completely in the claims.

Claims

1. A device comprising: a plurality of components; a memory including a keyword recognition module and a context recognition module; a microphone configured to receive an input speech spoken by a user; an analog-to-digital converter configured to convert the input speech from an analog form to a digital form and generate a digitized speech; a processor configured to: detect, using the keyword recognition module, a keyword in the digitized speech based on features extracted from the digitized speech; initiate, in response to detecting the keyword by the keyword recognition module, an action to be taken by one of the plurality of components, wherein the keyword is defined as a command to take the action; determine, using the context recognition module and after detecting the keyword, a context for the keyword based on additional features extracted from a portion of the digitized speech before and/or after the keyword, wherein the context is used to determine whether or not the keyword should be considered to be the command; and execute the action if the context determined by the context recognition module indicates that the keyword should be considered to be the command. 11. A method for speech recognition by a device having a microphone, a processor, and a memory including a keyword recognition module and a context recognition module, the method comprising: detecting, using the keyword recognition module, a keyword in a digitized speech based on features extracted from the digitized speech; initiating, in response to detecting the keyword by the keyword recognition module, an action to be taken by one of the plurality of components, wherein the keyword is defined as a command to take the action; determining, using the context recognition module and after detecting the keyword, a context for the keyword based on additional features extracted from a portion of the digitized speech before and/or after the keyword, wherein the context is used to determine whether or not the keyword should be considered to be the command; and executing the action if the context determined by the context recognition module indicates that the keyword should be considered to be the command. 20. A device comprising: a plurality of components; a memory including a keyword recognition module and a context recognition module; a microphone configured to receive an input speech spoken by a user; an analog-to-digital converter configured to convert the input speech from an analog form to a digital form and generate a digitized speech; a processor configured to: detect, using the keyword recognition module, a keyword in the digitized speech based on features extracted from the digitized speech; initiate, in response to detecting the keyword by the keyword recognition module, an action to be taken by one of the plurality of components, wherein the keyword is defined as a command to take the action; determine, using the context recognition module and after detecting the keyword, a context for the keyword based on additional features extracted from the digitized speech before and after the keyword, wherein the context is used to determine whether or not the keyword should be considered to be the command; execute the action if the context determined by the context recognition module indicates that the keyword should be considered to be the command; and terminate the action if the context determined by the context recognition module indicates that the keyword should not be considered to be the command.