WO2020140686A1 - 一种基于人脸检测唤醒智能设备的方法、装置及设备 - Google Patents

一种基于人脸检测唤醒智能设备的方法、装置及设备 Download PDF

Info

Publication number
WO2020140686A1
WO2020140686A1 PCT/CN2019/123351 CN2019123351W WO2020140686A1 WO 2020140686 A1 WO2020140686 A1 WO 2020140686A1 CN 2019123351 W CN2019123351 W CN 2019123351W WO 2020140686 A1 WO2020140686 A1 WO 2020140686A1
Authority
WO
WIPO (PCT)
Prior art keywords
face
graphics
smart device
awakening
face image
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/CN2019/123351
Other languages
English (en)
French (fr)
Inventor
鲁亚然
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Alibaba Group Holding Ltd
Original Assignee
Alibaba Group Holding Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Alibaba Group Holding Ltd filed Critical Alibaba Group Holding Ltd
Publication of WO2020140686A1 publication Critical patent/WO2020140686A1/zh
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Images

Classifications

    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F9/00Arrangements for program control, e.g. control units
    • G06F9/06Arrangements for program control, e.g. control units using stored programs, i.e. using an internal store of processing equipment to receive or retain programs
    • G06F9/44Arrangements for executing specific programs
    • G06F9/4401Bootstrapping

Definitions

  • the result of waking up the smart device is determined.
  • a prompt message is issued.
  • the extraction module is used to extract features of the face image
  • the above-mentioned effective face graphics include at least one of blinking movement, mouth opening movement and frowning movement.
  • the above-mentioned effective face graphics further include: one of a complete face graphic or a majority of the face area sufficient for face detection.
  • the above determination module is specifically used to wake up the smart device if the face graphics in the face image are valid face graphics; if the face graphics in the face image are invalid face graphics, do not wake up the smart device .
  • An embodiment of this specification provides a device for waking up a smart device based on face detection, including:
  • At least one processor and,
  • the result of waking up the smart device is determined.
  • FIG. 1 is a schematic flowchart of a method for awakening a smart device based on face detection according to an embodiment of the present specification
  • FIG. 2 is a schematic diagram of a wake-up smart device based on face detection and subsequent operation flow provided by an embodiment of this specification;
  • FIG. 3 is a schematic structural diagram of an apparatus for awakening a smart device based on face detection according to an embodiment of the present specification
  • FIG. 4 is another schematic structural diagram of an apparatus for waking up an intelligent device based on face detection provided by an embodiment of the present specification
  • FIG. 5 is a schematic structural diagram of a device for waking up a smart device based on face detection provided by an embodiment of the present specification.
  • Embodiments of the present specification provide a method, device, and device for waking up a smart device based on face detection.
  • Step 105 Collect a face image
  • Step 110 Extract features of the face image
  • the above-mentioned effective face graphics include at least one of a blinking motion, a mouth opening motion, and a frowning motion.
  • the above-mentioned effective face graphics further include one of a complete face graphic or a majority of the face area sufficient for face detection. It should be noted here that the face graphics in the face image can be rotated left and right around the neck. In the embodiment of the present specification, preferably, the face pattern in the face image whose rotation angle is within the range of 60 degrees to the left and right is determined as the majority of the face area sufficient for face detection.
  • the above-mentioned invalid face graphics include one of no face graphics or a face region that is insufficient for face detection. It should be noted here that the face pattern in the face image whose rotation angle is not within the range of 60 degrees to the left and right is determined as a face area that is insufficient for face detection.
  • Step 120 Determine the result of waking up the smart device according to the classification result
  • the “Tmall Genie” smart speaker wakes up and sends out a prompt message.
  • the above prompt information includes, but is not limited to, voice prompts, such as voice announcement "wake up”, light prompts, such as red light and/or vibration prompts, such as at least one of three slight vibrations.
  • whether the user is looking at the “Tmall Genie” smart speaker through iris detection is determined.
  • distance measurement is performed by a rangefinder to determine whether the distance between the user's face and the "Tmall Genie” smart speaker is within a certain range.
  • any interaction between the user and the "Tmall Genie” smart speaker needs to say “Tmall Genie", such as “Tmall Genie, play Zhou Huajian's song friends” in order to wake up the device to play music.
  • the technical solution only needs to say “play Zhou Huajian's song friend” to the "Tmall Genie” smart speaker to play the song “friend", making the human-computer interaction experience more friendly.
  • the "Tmall Genie” smart speaker responds to the detected face, such as emitting red light. After the user sees the red light, the "Tmall Genie” smart speaker voice inputs the user experience that he needs, such as "Play the news simulcast last night.” As shown in Figure 2, the "Tmall Genie” smart speaker searches for the news hookup last night based on the input voice information. As shown in Figure 2, the "Tmall Genie” smart speaker feedback information, such as the start of last night's news hookup.
  • FIG. 3 is a schematic structural diagram of an apparatus for awakening an intelligent device based on face detection provided by an embodiment of the present specification.
  • the structural schematic diagram includes: an acquisition module 305, an extraction module 310, a classification module 315, and a determination module 320;
  • the collection module 305 is used to collect face images
  • the extraction module 310 is used to extract features of the face image
  • the classification module 315 is used to classify the features of the face image through the trained classifier; wherein, the classification result includes that the face image contains valid face graphics or the face image contains invalid face graphics;
  • the extraction module 310 is specifically configured to perform convolution processing on the face image to obtain face features.
  • the effective face graphics include at least one of blinking movement, mouth opening movement, and frowning movement.
  • the invalid face graphics include: one of the face regions without face graphics or not enough for face detection.
  • the determination module 320 is specifically configured to wake up the smart device if the face graphics in the face image are valid face graphics; if the face graphics in the face image are invalid face graphics, do not wake up smart device.
  • the embodiment of the present specification provides another structural schematic diagram of an apparatus for waking up a smart device based on face detection. As shown in FIG. 4, compared to the structural schematic diagram shown in FIG. After the smart device is awakened, a prompt message is issued.
  • At least one processor 505 and,
  • a memory 510 communicatively connected to the at least one processor; wherein,
  • the memory stores instructions executable by the at least one processor, and the instructions are executed by the at least one processor to enable the at least one processor to:
  • the result of waking up the smart device is determined.
  • the embodiments of the present application may be provided as methods, systems, or computer program products. Therefore, the present invention may take the form of an entirely hardware embodiment, an entirely software embodiment, or an embodiment combining software and hardware. Moreover, the present application may take the form of a computer program product implemented on one or more computer usable storage media (including but not limited to disk storage, CD-ROM, optical storage, etc.) containing computer usable program code.
  • computer usable storage media including but not limited to disk storage, CD-ROM, optical storage, etc.
  • These computer program instructions may also be stored in a computer readable memory that can guide a computer or other programmable data processing device to work in a specific manner, so that the instructions stored in the computer readable memory produce an article of manufacture including an instruction device, the instructions
  • the device implements the functions specified in one block or multiple blocks of the flowchart one flow or multiple flows and/or block diagrams.
  • These computer program instructions can also be loaded onto a computer or other programmable data processing device, so that a series of operating steps are performed on the computer or other programmable device to generate computer-implemented processing, which is executed on the computer or other programmable device
  • the instructions provide steps for implementing the functions specified in one block or multiple blocks of the flowchart one flow or multiple flows and/or block diagrams.
  • the computing device includes one or more processors (CPUs), input/output interfaces, network interfaces, and memory.
  • processors CPUs
  • input/output interfaces network interfaces
  • memory volatile and non-volatile memory
  • the memory may include non-permanent memory, random access memory (RAM) and/or non-volatile memory in a computer-readable medium, such as read only memory (ROM) or flash memory (flash RAM). Memory is an example of computer-readable media.
  • RAM random access memory
  • ROM read only memory
  • flash RAM flash memory
  • Computer-readable media including permanent and non-permanent, removable and non-removable media, can store information by any method or technology.
  • the information may be computer readable instructions, data structures, modules of programs, or other data.
  • Examples of computer storage media include, but are not limited to, phase change memory (PRAM), static random access memory (SRAM), dynamic random access memory (DRAM), other types of random access memory (RAM), read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory or other memory technologies, read-only compact disc read-only memory (CD-ROM), digital versatile disc (DVD) or other optical storage, Magnetic tape cassettes, magnetic tape magnetic disk storage or other magnetic storage devices or any other non-transmission media can be used to store information that can be accessed by computing devices.
  • computer-readable media does not include temporary computer-readable media (transitory media), such as modulated data signals and carrier waves.

Landscapes

  • Engineering & Computer Science (AREA)
  • Software Systems (AREA)
  • Theoretical Computer Science (AREA)
  • Computer Security & Cryptography (AREA)
  • Physics & Mathematics (AREA)
  • General Engineering & Computer Science (AREA)
  • General Physics & Mathematics (AREA)
  • Processing Or Creating Images (AREA)
  • Image Analysis (AREA)

Abstract

一种基于人脸检测唤醒智能设备的方法、装置及设备。该方法包括:采集人脸图像(105);提取所述人脸图像的特征(110);通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形(115);根据分类结果,确定唤醒智能设备的结果(120)。

Description

一种基于人脸检测唤醒智能设备的方法、装置及设备 技术领域
本说明书涉及计算机技术领域,尤其是涉及一种基于人脸检测唤醒智能设备的方法、装置及设备。
背景技术
人机交互技术的蓬勃发展给人们带来了便捷的生活。日常生活中,越来越多用户开始使用智能设备方便自己的生活。
现有的人机交互流程一般可分为五个环节,包括:唤醒、响应、输入、理解、反馈。其中唤醒作为用户跟智能设备交互的第一个环节,尤为重要。如“天猫精灵”智能音箱接收到用户的语音输入“天猫精灵”这一唤醒词后会被唤醒。“天猫精灵”智能音箱被唤醒后,用户才可以利用这一智能设备进行音乐欣赏、听新闻等等。若“天猫精灵”智能音箱接收到用户的语音输入“播放一首歌曲”,即会播放一首歌曲;若“天猫精灵”智能音箱接收到用户的语音输入“我想听新闻”,即会播放当天新闻。然而,用户使用“天猫精灵”智能音箱进行音乐欣赏、听新闻等之前,需要语音输入唤醒词“天猫精灵”。因唤醒词是被限定之后的词语,使得用户与“天猫精灵”智能音箱之间的交互体验较差,显得及其不友好。
发明内容
本说明书实施例提供一种基于人脸检测唤醒智能设备的方法、装置及设备。解决了人机交互流程中需要通过唤醒词唤醒智能设备的问题。
为解决上述技术问题,本说明书实施例是这样实现的:
本说明书实施例提供的一种基于人脸检测唤醒智能设备的方法,该方法包括:
采集人脸图像;
提取所述人脸图像的特征;
通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形;
根据分类结果,确定唤醒智能设备的结果。
优选地,上述提取所述人脸图像的特征,包括:对所述人脸图像进行卷积处理,获得人脸特征。
优选地,上述有效人脸图形包括:眨眼动作、张嘴动作、皱眉动作中的至少一种。
优选地,上述有效人脸图形还包括:完整的人脸图形或足够进行人脸检测的大部分人脸区域中的一种。
优选地,上述无效人脸图形包括:没有人脸图形或不足够进行人脸检测的人脸区域中的一种。
优选地,上述根据分类结果,确定唤醒智能设备的结果,包括:
若人脸图像中人脸图形为有效人脸图形,则唤醒智能设备;
若人脸图像中人脸图形为无效人脸图形,则不唤醒智能设备。
优选地,上述智能设备被唤醒后,发出提示信息。
本说明书实施例提供的一种基于人脸检测唤醒智能设备的装置,该装置包括:采集模块、提取模块、分类模块和确定模块;
所述采集模块,用于采集人脸图像;
所述提取模块,用于提取所述人脸图像的特征;
所述分类模块,用于通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形;
所述确定模块,用于根据分类结果,确定唤醒智能设备的结果。
优选地,上述提取模块,具体用于对所述人脸图像进行卷积处理,获得人脸特征。
优选地,上述所述有效人脸图形包括:眨眼动作、张嘴动作、皱眉动作中的至少一种。
优选地,上述有效人脸图形还包括:完整的人脸图形或足够进行人脸检测的大部分人脸区域中的一种。
优选地,上述无效人脸图形包括:没有人脸图形或不足够进行人脸检测的人脸区域中的一种。
优选地,上述确定模块,具体用于若人脸图像中的人脸图形为有效人脸图形,则唤醒智能设备;若人脸图像中的人脸图形为无效人脸图形,则不唤醒智能设备。
优选地,上述装置还包括发出模块,用于所述智能设备被唤醒后,发出提示信息。
本说明书实施例提供的一种基于人脸检测唤醒智能设备的设备,包括:
至少一个处理器;以及,
与所述至少一个处理器通信连接的存储器;其中,
所述存储器存储有可被所述至少一个处理器执行的指令,所述指令被所述至少一个处理器执行,以使所述至少一个处理器能够:
采集人脸图像;
提取所述人脸图像的特征;
通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形;
根据分类结果,确定唤醒智能设备的结果。
本说明书实施例采用的上述至少一个技术方案能够达到以下有益效果:与现有技术中基于唤醒词唤醒智能设备的人机交互方式相比,本技术方案中通过检测到人脸来唤醒智能设备,使得用户与智能设备之间的交互体验更加友好。
附图说明
为了更清楚地说明本说明书实施例中的技术方案,下面将对实施例描述中所需要使用的附图作简单地介绍,显而易见地,下面描述中的附图仅仅是本说明书中记载的一些实施例,对于本领域普通技术人员来讲,在不付出创造性劳动的前提下,还可以根据这些附图获得其他的附图。
图1为本说明书实施例提供的一种基于人脸检测唤醒智能设备的方法的流程示意图;
图2本说明书实施例提供的基于人脸检测唤醒智能设备及后续操作流程示意图;
图3为本说明书实施例提供的一种基于人脸检测唤醒智能设备的装置的结构示意图;
图4为本说明书实施例提供的一种基于人脸检测唤醒智能设备的装置的另一结构示意图;
图5为本说明书实施例提供的一种基于人脸检测唤醒智能设备的设备的结构示意图。
具体实施方式
本说明书实施例提供一种基于人脸检测唤醒智能设备的方法、装置以及设备。
为了使本技术领域的人员更好地理解本说明书中的技术方案,下面将结合本说明书实施例中的附图,对本说明书实施例中的技术方案进行清楚、完整地描述,显然,所描述的实施例仅仅是本申请一部分实施例,而不是全部的实施例。基于本说明书实施例,本领域普通技术人员在没有作出创造性劳动前提下所获得的所有其他实施例,都应当属于本申请保护的范围。
现如今,智能设备在人们的日常生活中越来越普及。在一些应用场景中,要求用户对智能设备说出唤醒词唤醒智能设备后用户才能与智能设备进行后续的人机交互流程。通常,使用唤醒词作为唤醒智能设备的方式,使得人机交互体验较差,并不友好,为解决现有技术中的上述问题,本说明书实施例提供了一种基于人脸检测唤醒智能设备的方法,如图1所示为该方法的流程示意图,该流程示意图包括:
步骤105,采集人脸图像;
在本说明书实施例中,以唤醒“天猫精灵”智能音箱为例。
现有的“天猫精灵”智能音箱以“天猫精灵”作为唤醒词。“天猫精灵”智能音箱接收用户的语音输入“天猫精灵,播放一首音乐”。上述语音输入中包括唤醒词“天猫精灵”,“天猫精灵”智能音箱被唤醒,随后播放一首音乐。唤醒词被限定为“天猫精灵”使得用户唤醒“天猫精灵”智能设备的体验较差,并不友好。为改进现状,在本说明书实施例中,通过检测人脸代替唤醒词“天猫精灵”。本技术方案首先对人脸图像进行采集。
步骤110,提取所述人脸图像的特征;
在本说明书实施例中,作为一种可选地实施方式,执行步骤110之前,选取大量训练样本用于训练卷积神经网络。利用训练好的卷积神经网络卷积层对步骤105中采集到的人脸图像进行卷积处理,获得人脸特征。上述人脸特征为一组特征,用一个向量表 示,向量的每个维度代表一个特征,每个特征用一个特征值表示。
步骤115,通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形;
在本说明书实施例中,作为一种可选地实施方式,在执行步骤115之前,需要训练分类器。上述分类器包括但不限于逻辑回归、决策树、朴素贝叶斯、随机森林、GBDT(Gradient Boosting Decision Tree,梯度提升树)、深度神经网络中的一种。在训练分类器时,可以将含有有效人脸图形的人脸图像正样本的特征和不含有有效人脸图形的人脸图像负样本的特征分别赋予正样本标签和负样本标签,并将其输入到分类器中进行训练。当需要被判断的目标人脸图像的特征被输入到分类器时,分类器将目标人脸图像判定为含有有效人脸图形或将目标人脸图像判定为含有无效人脸图形。在本说明书实施例中,作为一种可选地实施方式,上述有效人脸图形包括眨眼动作、张嘴动作、皱眉动作中的至少一种。作为另一种可选地实施方式,上述有效人脸图形还包括完整的人脸图形或足够进行人脸检测的大部分人脸区域中的一种。在此需要说明的是,人脸图像中的人脸图形可以以脖子为支点左右转动。在本说明书实施例中,优选地,将人脸图像中的人脸图形转动角度在左右60度范围内的人脸图形确定为足够进行人脸检测的大部分人脸区域。作为一种可选地实施方式,上述无效人脸图形包括没有人脸图形或不足够进行人脸检测的人脸区域中的一种。在此需要说明的是,将人脸图像中的人脸图形转动角度不在左右60度范围内的人脸图形确定为不足够进行人脸检测的人脸区域。
步骤120,根据分类结果,确定唤醒智能设备的结果;
在本说明书实施例中,作为一种可选地实施方式,若步骤115中的分类结果人脸图像中人脸图形为有效人脸图形,则唤醒“天猫精灵”智能音箱;若步骤115中的分类结果人脸图像中人脸图形为无效人脸图形,则不唤醒“天猫精灵”智能音箱;
在本说明书实施例中,作为一种可选地实施方式,为进一步提高人机交互体验,“天猫精灵”智能音箱被唤醒后,发出提示信息。上述提示信息包括但不限于语音提示,如语音播报“已经睡醒”、灯光提示,如发出红光和/或振动提示,如轻微振动三次中的至少一种。
在此需要说明的是,本说明书实施例提供了如图2所示的基于人脸检测唤醒智能设备及后续操作流程示意图以说明“天猫精灵”智能音箱被唤醒后,用户与该设备的人机交互后续流程。如图2所示,当人脸满足预设的条件时,如用户双目注视“天猫精灵” 智能音箱、或用户面部距离“天猫精灵”智能音箱的距离在一定范围内或其他条件,“天猫精灵”智能音箱检测到人脸,确定为用户要与“天猫精灵”智能音箱进行沟通。作为一种可选地实施方式,通过虹膜检测判断用户双目是否注视“天猫精灵”智能音箱。作为一种可选地实施方式,通过测距仪进行距离测量,判断用户面部距离“天猫精灵”智能音箱的距离是否在一定范围内。对比之前用户与“天猫精灵”智能音箱的任何交互都需要说“天猫精灵”,比如“天猫精灵,播放周华健的歌曲朋友”,才能唤醒设备播放音乐。本技术方案只需对着“天猫精灵”智能音箱说“播放周华健的歌曲朋友”即可播放歌曲“朋友”,使人机交互体验更加友好。为进一步提高用户体验,如图2所示,“天猫精灵”智能音箱响应检测到的人脸,如发出红光。用户看到红光后,“天猫精灵”智能音箱语音输入自己需要的用户体验,如“播放昨晚的新闻联播”。如图2所示,“天猫精灵”智能音箱根据输入的语音信息搜索到昨晚的新闻联播。如图2所示,“天猫精灵”智能音箱反馈信息,如开始播放昨晚的新闻联播。
与基于唤醒词唤醒智能设备的人机交互方式相比,本说明书实施例采用的上述技术方案能够达到以下有益效果:本技术方案中通过检测到人脸来唤醒智能设备,使得用户与智能设备之间的交互体验更加友好。
图3为本说明书实施例提供的一种基于人脸检测唤醒智能设备的装置的结构示意图,该结构示意图包括:采集模块305、提取模块310、分类模块315和确定模块320;
所述采集模块305,用于采集人脸图像;
所述提取模块310,用于提取所述人脸图像的特征;
所述分类模块315,用于通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形;
所述确定模块320,用于根据分类结果,确定唤醒智能设备的结果。
优选地,所述提取模块310,具体用于对所述人脸图像进行卷积处理,获得人脸特征。
优选地,所述有效人脸图形包括:眨眼动作、张嘴动作、皱眉动作中的至少一种。
优选地,所述有效人脸图形还包括:完整的人脸图形或足够进行人脸检测的大部分人脸区域中的一种。
优选地,所述无效人脸图形包括:没有人脸图形或不足够进行人脸检测的人脸区 域中的一种。
优选地,所述确定模块320,具体用于若人脸图像中的人脸图形为有效人脸图形,则唤醒智能设备;若人脸图像中的人脸图形为无效人脸图形,则不唤醒智能设备。
优选地,为进一步使得人机交互流程中的唤醒流程更加人性化,更加友好。本说明书实施例提供了一种基于人脸检测唤醒智能设备的装置的另一结构示意图,如图4所示,相较于图3所示的结构示意图,该结构示意图增加了发出模块405,用于所述智能设备被唤醒后,发出提示信息。
图5为本说明书实施例提供的一种基于人脸检测唤醒智能设备的设备,包括:
至少一个处理器505;以及,
与所述至少一个处理器通信连接的存储器510;其中,
所述存储器存储有可被所述至少一个处理器执行的指令,所述指令被所述至少一个处理器执行,以使所述至少一个处理器能够:
采集人脸图像;
提取所述人脸图像的特征;
通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形;
根据分类结果,确定唤醒智能设备的结果。
本领域内的技术人员应明白,本申请的实施例可提供为方法、系统、或计算机程序产品。因此,本发明可采用完全硬件实施例、完全软件实施例、或结合软件和硬件方面的实施例的形式。而且,本申请可采用在一个或多个其中包含有计算机可用程序代码的计算机可用存储介质(包括但不限于磁盘存储器、CD-ROM、光学存储器等)上实施的计算机程序产品的形式。
本申请是参照根据本申请实施例的方法、设备(系统)、和计算机程序产品的流程图和/或方框图来描述的。应理解可由计算机程序指令实现流程图和/或方框图中的每一流程和/或方框、以及流程图和/或方框图中的流程和/或方框的结合。可提供这些计算机程序指令到计算机、专用计算机、嵌入式处理机或其他可编程数据处理设备的处理器以产生一个机器,使得通过计算机或其他可编程数据处理设备的处理器执行的指令产生用于实现在流程图一个流程或多个流程和/或方框图一个方框或多个方框中指 定的功能的装置。
这些计算机程序指令也可存储在能引导计算机或其他可编程数据处理设备以特定方式工作的计算机可读存储器中,使得存储在该计算机可读存储器中的指令产生包括指令装置的制造品,该指令装置实现在流程图一个流程或多个流程和/或方框图一个方框或多个方框中指定的功能。
这些计算机程序指令也可装载到计算机或其他可编程数据处理设备上,使得在计算机或其他可编程设备上执行一系列操作步骤以产生计算机实现的处理,从而在计算机或其他可编程设备上执行的指令提供用于实现在流程图一个流程或多个流程和/或方框图一个方框或多个方框中指定的功能的步骤。
在一个典型的配置中,计算设备包括一个或多个处理器(CPU)、输入/输出接口、网络接口和内存。
内存可能包括计算机可读介质中的非永久性存储器,随机存取存储器(RAM)和/或非易失性内存等形式,如只读存储器(ROM)或闪存(flash RAM)。内存是计算机可读介质的示例。
计算机可读介质包括永久性和非永久性、可移动和非可移动媒体可以由任何方法或技术来实现信息存储。信息可以是计算机可读指令、数据结构、程序的模块或其他数据。计算机的存储介质的例子包括,但不限于相变内存(PRAM)、静态随机存取存储器(SRAM)、动态随机存取存储器(DRAM)、其他类型的随机存取存储器(RAM)、只读存储器(ROM)、电可擦除可编程只读存储器(EEPROM)、快闪记忆体或其他内存技术、只读光盘只读存储器(CD-ROM)、数字多功能光盘(DVD)或其他光学存储、磁盒式磁带,磁带磁磁盘存储或其他磁性存储设备或任何其他非传输介质,可用于存储可以被计算设备访问的信息。按照本文中的界定,计算机可读介质不包括暂存电脑可读媒体(transitory media),如调制的数据信号和载波。
还需要说明的是,术语“包括”、“包含”或者其任何其他变体意在涵盖非排他性的包含,从而使得包括一系列要素的过程、方法、商品或者设备不仅包括那些要素,而且还包括没有明确列出的其他要素,或者是还包括为这种过程、方法、商品或者设备所固有的要素。在没有更多限制的情况下,由语句“包括一个……”限定的要素,并不排除在包括要素的过程、方法、商品或者设备中还存在另外的相同要素。
以上仅为本说明书的实施例而已,并不用于限制本说明书。对于本领域技术人员 来说,本说明书可以有各种更改和变化。凡在本说明书的精神和原理之内所作的任何修改、等同替换、改进等,均应包含在本说明书的权利要求范围之内。

Claims (15)

  1. 一种基于人脸检测唤醒智能设备的方法,其特征在于,该方法包括:
    采集人脸图像;
    提取所述人脸图像的特征;
    通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形;
    根据分类结果,确定唤醒智能设备的结果。
  2. 根据权利要求1所述的基于人脸检测唤醒智能设备的方法,其特征在于,所述提取所述人脸图像的特征,包括:对所述人脸图像进行卷积处理,获得人脸特征。
  3. 根据权利要求1所述的基于人脸检测唤醒智能设备的方法,其特征在于,所述有效人脸图形包括:眨眼动作、张嘴动作、皱眉动作中的至少一种。
  4. 根据权利要求3所述的基于人脸检测唤醒智能设备的方法,其特征在于,所述有效人脸图形还包括:完整的人脸图形或足够进行人脸检测的大部分人脸区域中的一种。
  5. 根据权利要求4所述的基于人脸检测唤醒智能设备的方法,其特征在于,所述无效人脸图形包括:没有人脸图形或不足够进行人脸检测的人脸区域中的一种。
  6. 根据权利要求5所述的基于人脸检测唤醒智能设备的方法,其特征在于,所述根据分类结果,确定唤醒智能设备的结果,包括:
    若人脸图像中人脸图形为有效人脸图形,则唤醒智能设备;
    若人脸图像中人脸图形为无效人脸图形,则不唤醒智能设备。
  7. 根据权利要求6所述的基于人脸检测唤醒智能设备的方法,其特征在于,所述智能设备被唤醒后,发出提示信息。
  8. 一种基于人脸检测唤醒智能设备的装置,其特征在于,该装置包括:采集模块、提取模块、分类模块和确定模块;
    所述采集模块,用于采集人脸图像;
    所述提取模块,用于提取所述人脸图像的特征;
    所述分类模块,用于通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形;
    所述确定模块,用于根据分类结果,确定唤醒智能设备的结果。
  9. 根据权利要求8所述的基于人脸检测唤醒智能设备的装置,其特征在于,所述提取模块,具体用于对所述人脸图像进行卷积处理,获得人脸特征。
  10. 根据权利要求8所述的基于人脸检测唤醒智能设备的装置,其特征在于,所述 有效人脸图形包括:眨眼动作、张嘴动作、皱眉动作中的至少一种。
  11. 根据权利要求10所述的基于人脸检测唤醒智能设备的装置,其特征在于,所述有效人脸图形还包括:完整的人脸图形或足够进行人脸检测的大部分人脸区域中的一种。
  12. 根据权利要求11所述的基于人脸检测唤醒智能设备的装置,其特征在于,所述无效人脸图形包括:没有人脸图形或不足够进行人脸检测的人脸区域中的一种。
  13. 根据权利要求12所述的基于人脸检测唤醒智能设备的装置,其特征在于,所述确定模块,具体用于若人脸图像中的人脸图形为有效人脸图形,则唤醒智能设备;若人脸图像中的人脸图形为无效人脸图形,则不唤醒智能设备。
  14. 根据权利要求13所述的基于人脸检测唤醒智能设备的装置,其特征在于,所述装置还包括发出模块,用于所述智能设备被唤醒后,发出提示信息。
  15. 一种基于人脸检测唤醒智能设备的设备,包括:
    至少一个处理器;以及,
    与所述至少一个处理器通信连接的存储器;其中,
    所述存储器存储有可被所述至少一个处理器执行的指令,所述指令被所述至少一个处理器执行,以使所述至少一个处理器能够:
    采集人脸图像;
    提取所述人脸图像的特征;
    通过训练后的分类器对所述人脸图像的特征进行分类;其中,分类结果包括人脸图像中含有有效人脸图形或人脸图像中含有无效人脸图形;
    根据分类结果,确定唤醒智能设备的结果。
PCT/CN2019/123351 2019-01-03 2019-12-05 一种基于人脸检测唤醒智能设备的方法、装置及设备 Ceased WO2020140686A1 (zh)

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
CN201910004948.7 2019-01-03
CN201910004948.7A CN109725946A (zh) 2019-01-03 2019-01-03 一种基于人脸检测唤醒智能设备的方法、装置及设备

Publications (1)

Publication Number Publication Date
WO2020140686A1 true WO2020140686A1 (zh) 2020-07-09

Family

ID=66298125

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/CN2019/123351 Ceased WO2020140686A1 (zh) 2019-01-03 2019-12-05 一种基于人脸检测唤醒智能设备的方法、装置及设备

Country Status (2)

Country Link
CN (1) CN109725946A (zh)
WO (1) WO2020140686A1 (zh)

Citations (5)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US20070201730A1 (en) * 2006-02-20 2007-08-30 Funai Electric Co., Ltd. Television set and authentication device
US20160171285A1 (en) * 2014-12-12 2016-06-16 Samsung Electronics Co., Ltd. Method of detecting object in image and image processing device
CN108363999A (zh) * 2018-03-22 2018-08-03 百度在线网络技术(北京)有限公司 基于人脸识别的操作执行方法和装置
CN108521516A (zh) * 2018-03-30 2018-09-11 百度在线网络技术(北京)有限公司 用于终端设备的控制方法和装置
CN108804893A (zh) * 2018-03-30 2018-11-13 百度在线网络技术(北京)有限公司 一种基于人脸识别的控制方法、装置和服务器

Family Cites Families (5)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
CN107102540A (zh) * 2016-02-23 2017-08-29 芋头科技(杭州)有限公司 一种唤醒智能机器人的方法及智能机器人
CN107798282B (zh) * 2016-09-07 2021-12-31 北京眼神科技有限公司 一种活体人脸的检测方法和装置
US10816800B2 (en) * 2016-12-23 2020-10-27 Samsung Electronics Co., Ltd. Electronic device and method of controlling the same
CN108664782B (zh) * 2017-03-28 2023-09-12 三星电子株式会社 面部验证方法和设备
CN108536027B (zh) * 2018-03-30 2020-11-03 百度在线网络技术(北京)有限公司 智能家居控制方法、装置和服务器

Patent Citations (5)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US20070201730A1 (en) * 2006-02-20 2007-08-30 Funai Electric Co., Ltd. Television set and authentication device
US20160171285A1 (en) * 2014-12-12 2016-06-16 Samsung Electronics Co., Ltd. Method of detecting object in image and image processing device
CN108363999A (zh) * 2018-03-22 2018-08-03 百度在线网络技术(北京)有限公司 基于人脸识别的操作执行方法和装置
CN108521516A (zh) * 2018-03-30 2018-09-11 百度在线网络技术(北京)有限公司 用于终端设备的控制方法和装置
CN108804893A (zh) * 2018-03-30 2018-11-13 百度在线网络技术(北京)有限公司 一种基于人脸识别的控制方法、装置和服务器

Also Published As

Publication number Publication date
CN109725946A (zh) 2019-05-07

Similar Documents

Publication Publication Date Title
US11070644B1 (en) Resource grouped architecture for profile switching
WO2020114384A1 (zh) 一种语音交互方法和装置
CN109065044B (zh) 唤醒词识别方法、装置、电子设备及计算机可读存储介质
CN108986822A (zh) 语音识别方法、装置、电子设备及非暂态计算机存储介质
US11783805B1 (en) Voice user interface notification ordering
CN110827821A (zh) 一种语音交互装置、方法和计算机可读存储介质
CN111292734A (zh) 一种语音交互方法和装置
US11749267B2 (en) Adapting hotword recognition based on personalized negatives
CN110866090A (zh) 用于语音交互的方法、装置、电子设备和计算机存储介质
US12217751B2 (en) Digital signal processor-based continued conversation
US10880384B1 (en) Multi-tasking resource management
US11398226B1 (en) Complex natural language processing
US12525250B2 (en) Cascade architecture for noise-robust keyword spotting
WO2020114323A1 (zh) 一种用于个性化语音合成的方法和装置
CN113948076A (zh) 语音交互方法、设备和系统
US11335346B1 (en) Natural language understanding processing
CN108830059A (zh) 媒体访问的控制方法、装置及电子设备
US11900921B1 (en) Multi-device speech processing
CN112017662B (zh) 控制指令确定方法、装置、电子设备和存储介质
CN111754989B (zh) 一种语音误唤醒的规避方法及电子设备
WO2020140686A1 (zh) 一种基于人脸检测唤醒智能设备的方法、装置及设备
CN115599891B (zh) 一种确定异常对话数据方法、装置、设备及可读存储介质
US12499309B1 (en) Programmatically updating machine learning models
Feng et al. Sample dropout for audio scene classification using multi-scale dense connected convolutional neural network
US12267286B1 (en) Sharing of content

Legal Events

Date Code Title Description
121 Ep: the epo has been informed by wipo that ep was designated in this application

Ref document number: 19907608

Country of ref document: EP

Kind code of ref document: A1

NENP Non-entry into the national phase

Ref country code: DE

122 Ep: pct application non-entry in european phase

Ref document number: 19907608

Country of ref document: EP

Kind code of ref document: A1