EP4542542A1 - Sound box playing control method, sound box playing control apparatus, and storage medium - Google Patents

Sound box playing control method, sound box playing control apparatus, and storage medium Download PDF

Info

Publication number
EP4542542A1
EP4542542A1 EP22946317.9A EP22946317A EP4542542A1 EP 4542542 A1 EP4542542 A1 EP 4542542A1 EP 22946317 A EP22946317 A EP 22946317A EP 4542542 A1 EP4542542 A1 EP 4542542A1
Authority
EP
European Patent Office
Prior art keywords
speaker
sub
subspace
target
audio data
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Pending
Application number
EP22946317.9A
Other languages
German (de)
French (fr)
Other versions
EP4542542A4 (en
Inventor
Lingsong Zhou
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Beijing Xiaomi Mobile Software Co Ltd
Beijing Xiaomi Pinecone Electronic Co Ltd
Original Assignee
Beijing Xiaomi Mobile Software Co Ltd
Beijing Xiaomi Pinecone Electronic Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Beijing Xiaomi Mobile Software Co Ltd, Beijing Xiaomi Pinecone Electronic Co Ltd filed Critical Beijing Xiaomi Mobile Software Co Ltd
Publication of EP4542542A1 publication Critical patent/EP4542542A1/en
Publication of EP4542542A4 publication Critical patent/EP4542542A4/en
Pending legal-status Critical Current

Links

Images

Classifications

    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04SSTEREOPHONIC SYSTEMS 
    • H04S7/00Indicating arrangements; Control arrangements, e.g. balance control
    • H04S7/30Control circuits for electronic adaptation of the sound field
    • H04S7/302Electronic adaptation of stereophonic sound system to listener position or orientation
    • H04S7/303Tracking of listener position or orientation
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04RLOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
    • H04R3/00Circuits for transducers
    • H04R3/12Circuits for transducers for distributing signals to two or more loudspeakers
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04RLOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
    • H04R2227/00Details of public address [PA] systems covered by H04R27/00 but not provided for in any of its subgroups
    • H04R2227/003Digital PA systems using, e.g. LAN or internet
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04RLOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
    • H04R2227/00Details of public address [PA] systems covered by H04R27/00 but not provided for in any of its subgroups
    • H04R2227/005Audio distribution systems for home, i.e. multi-room use

Definitions

  • the present disclosure relates to the field of speakers, in particular to a speaker play control method, a speaker play control device and a storage medium.
  • the whole house speaker brings shocking auditory enjoyment to people's home life.
  • a plurality of smart speakers are connected to each other through WiFi and play music synchronously, so that users can have an immersive listening experience while walking.
  • the terminal speaker gets audio data from the server and transmits it to the speakers in all rooms through the wireless network. Through a certain synchronization mechanism, the same audio frame can be played at the same time.
  • the users usually only play music in one or part of the rooms, and playing music in the rooms without the users will lead to the loss of electric energy on the one hand.
  • the speakers sending the data will occupy the network channel, the more network resources are occupied, the more likely it will lead to instability of the system. For example, the problem such as unsynchronized play will occur.
  • the present disclosure provides a speaker play control method, a speaker play control device and a storage medium.
  • a speaker play control method which is applied to a master speaker, and the master speaker is configured to control a plurality of sub-speakers, and the plurality of sub-speakers belong to different subspaces of a same space.
  • the method includes: in response to enabling an audio play function, acquiring audio data from a server, and determining a target subspace in the subspaces to which the plurality of sub-speakers belong, wherein the target subspace is a subspace where a user is currently located; and sending the audio data to a sub-speaker in the target subspace, and controlling the sub-speaker in the target subspace to play audio based on the audio data.
  • the determining the target subspace in the subspaces to which the plurality of sub-speakers belong includes: determining a user detection result, wherein the user detection result is determined by a set sub-speaker in the subspace based on detection of a human body activity of the user, and the user detection result comprises a presence or an absence of the user; and determining a subspace to which a target sub-speaker whose detection result is the presence of the user belongs as the target subspace.
  • the plurality of sub-speakers correspond to spatial identifiers
  • the spatial identifier is configured to identify the subspace to which the sub-speaker belongs
  • the sending the audio data to the sub-speaker in the target subspace and controlling the sub-speaker in the target subspace to play audio based on the audio data includes: determining a target spatial identifier corresponding to the target subspace; and sending the audio data to the sub-speaker in the target subspace identified by the target spatial identifier, and controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data.
  • the target spatial identifier comprises a first target spatial identifier and a second target spatial identifier, a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves;
  • the controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data includes: sending a first control play instruction to the sub-speaker in the first target subspace, and controlling the sub-speaker in the first target subspace based on the first control play instruction to play audio based on the first control play instruction and the audio data, wherein the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and sending a second control play instruction to the sub-speaker in the second target subspace, and controlling the sub-speaker in the second target subspace based on the second control play instruction to play
  • the method further includes: determining a non-target subspace, wherein the non-target subspace is a subspace without the user to which each set sub-speaker belongs; and stopping sending the audio data to the non-target subspace.
  • a speaker play control method which is applied to a sub-speaker, and a subspace to which the sub-speaker belongs is a target subspace, and the target subspace is a subspace where a user is currently located.
  • the method includes: acquiring audio data sent by a master speaker; and playing audio based on the audio data.
  • the sub-speaker corresponds to a spatial identifier
  • the spatial identifier is configured to identify the subspace to which the sub-speaker belongs
  • the acquiring the audio data sent by the master speaker includes: acquiring the audio data sent by the master speaker based on the spatial identifier.
  • the playing audio based on the audio data includes: in response to that the spatial identifier is a first target spatial identifier, acquiring a first control play instruction sent by the master speaker, wherein a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and based on the first control play instruction, playing audio based on the audio data in a manner that the playing sound changes from large to small.
  • the playing audio based on the audio data includes: in response to that the spatial identifier is a second target spatial identifier, acquiring a second control play instruction sent by the master speaker, wherein a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves, and the second control play instruction is configured to control a playing sound of the sub-speaker to change from small to large; and based on the second control play instruction, playing audio based on the audio data in a manner that the playing sound changes from small to large.
  • the playing audio based on the audio data includes: calibrating a system clock between the sub-speaker and the master speaker; and based on the calibrated system clock and the audio data, playing audio synchronously with the master speaker.
  • a speaker play control device which includes: a determining unit configured to acquire audio data from a server in response to enabling an audio play function, and determining a target subspace in subspaces to which a plurality of sub-speakers belong, wherein the target subspace is a subspace where a user is currently located; and a playing unit configured to send the audio data to a sub-speaker in the target subspace and control the sub-speaker in the target subspace to play audio based on the audio data.
  • the plurality of sub-speakers correspond to spatial identifiers
  • the spatial identifier is configured to identify the subspace to which the sub-speaker belongs
  • the playing unit sends the audio data to the sub-speaker in the target subspace and controls the sub-speaker in the target subspace to play audio based on the audio data in a following manner: determining a target spatial identifier corresponding to the target subspace; and sending the audio data to the sub-speaker in the target subspace identified by the target spatial identifier, and controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data.
  • the target spatial identifier comprises a first target spatial identifier and a second target spatial identifier, wherein a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves;
  • the playing unit controls the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data in a following manner; sending a first control play instruction to the sub-speaker in the first target subspace, and controlling the sub-speaker in the first target subspace based on the first control play instruction to play audio based on the first control play instruction and the audio data, wherein the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and sending a second control play instruction to the sub-speaker in the second target subspace, and controlling the sub-speaker in the second target subspace based
  • the playing unit is further configured to: determine a non-target subspace, wherein the non-target subspace is a subspace without the user to which each set sub-speaker belongs; and stop sending the audio data to the non-target subspace.
  • a speaker play control device which includes: an acquiring unit configured to acquire audio data sent by a master speaker; and a playing unit configured to play audio based on the audio data.
  • the sub-speaker is a set sub-speaker, and the set sub-speaker is configured to detect whether there is the user in a subspace to which a target sub-speaker belongs, and the playing unit is further configured to: detect a human body activity, and determining a user detection result based on a human body activity detection result, wherein the user detection result comprises a presence or an absence of the user; and send the user detection result to the master speaker.
  • the sub-speaker corresponds to a spatial identifier
  • the spatial identifier is configured to identify the subspace to which the sub-speaker belongs
  • the acquiring unit acquires the audio data sent by the master speaker in a following way: acquiring the audio data sent by the master speaker based on the spatial identifier.
  • the playing unit plays audio based on the audio data in a following manner: in response to that the spatial identifier is a first target spatial identifier, acquiring a first control play instruction sent by the master speaker, wherein a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and based on the first control play instruction, playing audio based on the audio data in a manner that the playing sound changes from large to small.
  • the playing unit plays audio based on the audio data in a following manner: in response to that the spatial identifier is a second target spatial identifier, acquiring a second control play instruction sent by the master speaker, wherein a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves, and the second control play instruction is configured to control a playing sound of the sub-speaker to change from small to large; and based on the second control play instruction, playing audio based on the audio data in a manner that the playing sound changes from small to large.
  • the playing unit plays audio based on the audio data in a following manner: calibrating a system clock between the sub-speaker and the master speaker; and based on the calibrated system clock and the audio data, playing audio synchronously with the master speaker.
  • a speaker play control device which includes: a processor; and a memory for storing instructions executable by the processor.
  • the processor is configured to perform the method described in the first aspect or any one of the embodiments of the first aspect.
  • a speaker play control device which includes: a processor; and a memory for storing instructions executable the processor.
  • the processor is configured to perform the method described in the second aspect or any one of the embodiments of the second aspect.
  • a computer-readable storage medium in which instructions are stored.
  • the instructions in the storage medium are executed by a processor of a network device, the network device is enabled to perform the method described in the first aspect or any one of the embodiments of the first aspect.
  • a computer-readable storage medium in which instructions are stored.
  • the instructions in the storage medium are executed by a processor of a terminal, the terminal is enabled to perform the method described in the second aspect or any one of the embodiments of the second aspect.
  • the technical solution provided by the embodiments of the present disclosure can include the following beneficial effects.
  • the audio data is acquired from the server through the master speaker, and the space where the user is currently located is determined.
  • the master speaker sends the audio data acquired from the server to the sub-speaker in the space where the user is currently located, and the sub-speaker plays audio based on the audio data.
  • the speaker in the space selectively plays the audio data according to the space where the user is currently located, thus reducing the waste of energy. That is, in the current space, the speaker which does not detect that the user is in this space is in a non-working state.
  • the channel resources between the master speaker and the sub-speakers that do not play the audio data are released, thus providing the user with a more stable listening experience.
  • a speaker play control method provided by embodiments of the present disclosure may be applied to a plurality of smart devices connected by wireless communication technology, so as to realize an application scene of sharing music between users.
  • the plurality of smart devices are in a set space.
  • the plurality of smart devices belong to different subspaces respectively, and the plurality of smart devices include a master smart device and a sub-smart device.
  • the smart device may be a smart speaker or other smart devices.
  • the speakers are interconnected through a wireless network to realize the whole house play function.
  • a plurality of smart speakers are connected to each other through WiFi and play music synchronously, so that the user can have an immersive listening experience while walking.
  • the speakers get the music data from the server and send it to the speakers in all rooms through the wireless network.
  • the same audio frame can be played at the same time, so that all the smart speakers in the family can play music synchronously.
  • the audio data is processed respectively.
  • the master speaker and the other sub-speakers are set to play the same audio frame after a fixed time, so that the plurality of speakers can play the audio synchronously, thus achieving the listening experience of the whole house play.
  • the user usually only plays music in one or part of the rooms, and playing music in the room without the user will waste electric energy on the one hand.
  • sending the data will occupy the network channel. The more network resources are occupied, the more likely it will lead to instability of the system. For example, the plurality of speakers cannot play the audio data synchronously.
  • the present disclosure provides a speaker play control method.
  • the audio data is acquired from the server through the master speaker, and a target subspace is determined in subspaces to which a plurality of sub-speakers belong.
  • the target subspace is a subspace where the user is currently located, and the subspace is one of a plurality of spaces into which a complete space is divided.
  • the master speaker sends the audio data acquired from the server to the sub-speaker in the space where the user is currently located, and the sub-speaker plays audio based on the audio data.
  • the speaker in the space selectively plays the audio data according to the space where the user is currently located, thus reducing the waste of energy, that is, for a current space, when no user is detected in this space, the speaker in this space is in a non-working state.
  • the channel resources between the master speaker and the sub-speakers that do not play the audio data are released, which reduces the network load as much as possible, thus improving the stability of the system and providing the user with a more stable listening experience. Therefore, compared with the way of controlling the speaker to play in the related art, the speaker play control method provided by the present disclosure is more flexible and intelligent.
  • Fig. 1 is a flow chart of a speaker play control method according to an illustrative embodiment. As shown in Fig. 1 , the method is applied to a master speaker and includes the following steps.
  • a speaker communicating with the server and other speakers is called a master speaker, and a speaker communicating with the master speaker is called a sub-speaker.
  • the master speaker is used to control a plurality of sub-speakers, and different sub-speakers are included in different subspaces of the same space. One or more different sub-speakers may be included in the same subspace.
  • step S11 when it is determined that the master speaker is enabled with an audio play function, audio data is acquired from a server, and a target subspace is determined in subspaces to which a plurality of sub-speakers belong.
  • the target subspace is a subspace where the user is currently located.
  • the subspace may be understood as any room, such as a living room, a kitchen, a bedroom and so on.
  • the master speaker when it is determined that the user has enabled the audio play functions of all the speakers in the rooms, acquires the audio data from the server through a wireless transmission technology, and determines the room where the user is currently located, by receiving a signal sent by the sub-speaker in the subspace.
  • step S12 the audio data is sent to the sub-speaker in the target subspace, and the sub-speaker in the target subspace is controlled to play audio based on the audio data.
  • the master speaker determines which space has the user, it sends the audio data acquired from the server to the sub-speaker in the space where the user exists, and controls the sub-speaker in the space to play the decoded audio data.
  • the master speaker determines a non-target subspace through an instruction sent by the sub-speaker in the subspace, and the non-target subspace is a subspace to which a set sub-speaker whose detection result is the absence of the user belongs.
  • the master speaker determines the non-target subspace, it stops sending the audio data to the sub-speaker in the non-target subspace.
  • the master speaker in case that the master speaker determines that there is no user in the subspace, the master speaker stops sending the audio data to the sub-speaker in the space and closes a channel between the master speaker and the sub-speaker in the space.
  • the channel resources between the master speaker and the sub-speaker that does not play the audio data are released, which improves the stability of the system and provides the user with a more stable listening experience.
  • the audio data is acquired from the server, and the target subspace is determined in the subspaces to which the plurality of sub-speakers belong.
  • the audio data is sent to the sub-speaker in the target subspace, and the sub-speaker in the target subspace is controlled to play audio based on the audio data.
  • the speaker in the space selectively plays the audio data according to the space where the user is currently located, thereby reducing the waste of energy. It can be understood that only when there is the user in the space, the master speaker sends the audio data to the sub-speaker in the space.
  • Fig. 2 is a flow chart of determining a target subspace according to an illustrative embodiment. As shown in fig. 2 , determining the target subspace in the subspaces to which the plurality of sub-speakers belong includes the following steps.
  • step S21 a user detection result is determined.
  • the user detection result is determined by a set sub-speaker in the subspace based on the detection of the human body activity of the user, and the user detection result includes the presence or absence of the user.
  • step S22 the subspace to which a target sub-speaker whose detection result is the presence of the user belongs is determined as the target subspace.
  • the master speaker receives an instruction confirming the presence of the user from the sub-speaker, and determines the subspace with the user to which the target sub-speaker belongs as the target subspace. If the sub-speaker set in the subspace does not detect the human body activity of the user in the space, the master speaker receives an instruction for confirming the absence of the user from the sub-speaker.
  • the user detection result is determined, and the subspace to which the target sub-speaker whose detection result is the presence of the user belongs is determined as the target subspace.
  • the master speaker based on the detection result instruction transmitted from the sub-speaker to the master speaker, the master speaker sends a corresponding instruction to control the sub-speaker, thus realizing the interaction between the sub-speaker and the master speaker.
  • Fig. 3 is a flow chart of controlling a sub-speaker in a target subspace to play audio based on audio data according to an illustrative embodiment. As shown in Fig. 3 , sending the audio data to the sub-speaker in the target subspace and controlling the sub-speaker in the target subspace to play audio based on the audio data includes the following steps.
  • step S31 a target spatial identifier corresponding to the target subspace is determined.
  • the plurality of sub-speakers correspond to spatial identifiers, and the spatial identifier is used to identify the subspace to which the sub-speaker belongs.
  • the room environment is divided into a plurality of subspaces, such as a kitchen, a living room, a bedroom and other subspaces.
  • the user can select the space where the speaker is located. For example, the user can select the space where the speaker is currently located through a resizable list on a display screen of the speaker, so that each speaker knows the space environment where it is currently located. For example, speaker A and speaker B are currently placed in the living room, and speaker C and speaker D are currently placed in the master bedroom and so on.
  • the sub-speakers in the target subspace perceive each other through ultrasonic communication to determine the subspace to which each sub-speaker belongs.
  • the sub-speaker sends its own spatial identifier to the master speaker, and the master speaker stores IP addresses or other identification IDs of all the sub-speakers in the subspace in a list of devices, so that a target spatial identifier corresponding to the target subspace can be determined.
  • step S32 the audio data is sent to the sub-speaker in the target subspace identified by the target spatial identifier, and the sub-speaker in the target subspace are controlled based on the target spatial identifier to play audio based on the audio data.
  • the target spatial identifier corresponding to the target subspace is determined.
  • the audio data is sent to the sub-speaker in the target subspace identified by the target spatial identifier, and the sub-speaker in the target subspace is controlled based on the target spatial identifier to play audio based on the audio data.
  • the master speaker can accurately control the sub-speaker in the space where the user exists to play audio.
  • Fig. 4 is a flow chart of controlling a sub-speaker in a target subspace to play audio based on audio data according to an illustrative embodiment. As shown in Fig. 4 , controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data includes the following steps.
  • step S41 a first control play instruction is sent to a sub-speaker in a first target subspace, and the sub-speaker in the first target subspace is controlled based on the first control play instruction to play audio based on the first control play instruction and the audio data.
  • the first control play instruction is used to control a playing sound of the sub-speaker to change from large to small.
  • the sub-speaker in the living room cannot detect the user's human body activity at this time, and the sub-speaker sends the instruction for confirming the absence of the user to the master speaker.
  • the master speaker sends a volume reduction control instruction (i.e. the first control play instruction) to the sub-speaker in the living room, and controls the playing sound of the sub-speaker to change from large to small until there is no sound.
  • step S42 a second control play instruction is sent to the sub-speaker in a second target subspace, and the sub-speaker in the second target subspace is controlled to play audio based on the second control play instruction and the audio data.
  • the second control play instruction is used to control the playing sound of the sub-speaker to change from small to large.
  • the sub-speaker in the bedroom detects the human body activity of the user at this time, and the sub-speaker sends the instruction for confirming the presence of the user to the master speaker.
  • the master speaker sends a volume increase control instruction (i.e. the second control play instruction) to the sub-speaker in the bedroom to control the playing sound of the sub-speaker to change from small to large.
  • the first control play instruction is sent to the sub-speaker in the first target subspace, and the sub-speaker in the first target subspace is controlled based on the first control play instruction to play audio based on the first control play instruction and the audio data.
  • the second control play instruction is sent to the sub-speaker in the second target subspace, and the sub-speaker in the second target subspace is controlled based on the second control play instruction to play audio based on the second control play instruction and the audio data.
  • the master speaker can also control the change of volume of the audio played by the sub-speaker according to the detection result of the sub-speaker in the subspace, so that the sub-speaker can achieve a gradual effect in volume and bring a good listening experience to the user.
  • the audio data is stopped from being sent to the sub-speaker in the space.
  • the speaker in the space selectively plays the audio data according to the space where the user is currently located, thereby reducing the waste of energy.
  • the channel resources between the master speaker and the sub-speakers that do not play the audio data are released, which improves the stability of the system.
  • Fig. 5 is a flow chart of a speaker play control method according to an illustrative embodiment, which is applied to a sub-speaker as shown in Fig. 5 and includes the following steps.
  • step S51 the sub-speaker in the target subspace acquires the audio data sent by the master speaker.
  • step S52 audio is played based on the audio data.
  • the audio data sent by the master speaker is acquired, and the audio is played based on the audio data.
  • the sub-speaker is controlled by the master speaker, and the interaction between the sub-speaker and the master speaker is realized.
  • Fig. 6 is a flow chart of detecting whether a user exists in a subspace to which a target sub-speaker belongs according to an illustrative embodiment.
  • the sub-speaker is a set sub-speaker, and the set sub-speaker is used to detect whether the user exists in the subspace to which the target sub-speaker belongs, and the method includes the following steps.
  • step S61 the human body activity is detected, and the user detection result is determined based on a detection result of the human body activity.
  • the user detection results include the presence or absence of the user.
  • one sub-speaker is selected randomly or selected according to the performance of the sub-speaker in each subspace to transmit and receive an ultrasonic wave.
  • the selected sub-speaker in each subspace emits the ultrasonic wave to detect the human body activity of the user.
  • a differential channel impulse response (dCIR) is used to detect the human body activity of the user. The essence of this detection method is through calculating the dCIR in the environment in real-time. When the user is not active in subspace A at present, the dCIR value of the sub-speaker in subspace A receiving the ultrasonic wave back approaches zero.
  • the sub-speaker set in the target subspace will send the instruction that the user does not exist in the space to the master speaker.
  • the master speaker sends the corresponding instruction operation to the sub-speaker.
  • the dCIR value of the sub-speaker in subspace B receiving the ultrasonic wave back will reflect the amplitude change on the dCIR.
  • the speaker will send the instruction that the user exists in the space to the master speaker.
  • the master speaker sends the corresponding instruction operation to the sub-speaker. By detecting the overall amplitude state of the dCIR, it can be detected whether there is the user in the current subspace.
  • S is an ultrasonic signal emitted by a loudspeaker
  • r is the ultrasonic signal collected by a microphone
  • H is a CIR vector.
  • step S62 the user detection result is sent to the master speaker.
  • the instruction is sent to the master speaker, informing the master speaker that it can continue to send the audio frame data to the sub-speaker in the subspace.
  • the master speaker receives the instruction sent from the sub-speaker, it can be determined that the user detection result determined by the set sub-speaker in this subspace is that there is the user.
  • the master speaker determines the subspace to which the target speaker whose detection result is the presence of the user belongs as the target subspace.
  • the sub-speaker in the subspace may also detect that there is no user in the current subspace.
  • non-target subspace information is sent to the master speaker, and a non-target subspace is a subspace without the user to which each set sub-speaker belongs.
  • the master speaker is controlled to stop sending the audio data to the non-target subspace.
  • the sub-speaker when the dCIR value of the ultrasonic wave received back by the sub-speaker in the subspace approaches zero, it is determined that there is no user in the current subspace. When it is determined that there is no user in this subspace, the sub-speaker will send the instruction to the master speaker to inform that there is no user in the space to which it belongs.
  • the master speaker when the master speaker receives the non-target subspace information sent from the sub-speaker, it is determined that there is no user in the space where the sub-speaker is located. In order to avoid waste of resources, the master speaker stops sending the audio data to the non-target subspace.
  • the human body activity is detected, and the user detection result is determined based on the human body activity detection result.
  • the user detection result is sent to the master speaker.
  • the audio data is stopped in time from being sent to the sub-speaker in the subspace where the user does not exist, and the channel resources are released dynamically, thus ensuring the stability of the system and providing a better listening experience.
  • Fig. 7 is a flow chart of acquiring audio data sent by a master speaker according to an illustrative embodiment. As shown in Fig. 7 , the acquiring audio data sent by the master speaker includes the following steps.
  • step S71 the audio data sent by the master speaker based on the spatial identifier is acquired.
  • step S72 the audio is played based on different spatial identifiers.
  • the first control play instruction sent by the master speaker is acquired, and the first target subspace identified by the first target spatial identifier is the subspace where the user is before the user moves, and the first control play instruction is used to control the playing sound of the sub-speaker to change from large to small.
  • the audio is played based on the audio data in a manner that the playing sound changes from large to small.
  • the sub-speaker in the living room cannot detect the user's human body activity at this time, and the sub-speaker sends the instruction for confirming the absence of the user to the master speaker.
  • the master speaker sends the volume reduction control instruction (i.e. the first control play instruction) to the sub-speaker in the living room, and controls the playing sound of the sub-speaker to change from large to small until there is no sound.
  • the second control play instruction sent by the master speaker is acquired, and the second target subspace identified by the second target spatial identifier is the subspace where the user is after the user moves, and the second control play instruction is used to control the playing sound of the sub-speaker to change from small to large.
  • the audio is played based on the audio data in a manner that the playing sound changes from small to large.
  • the sub-speaker in the bedroom detects the user's human body activity at this time, and the sub-speaker sends the instruction for confirming the presence of the user to the master speaker.
  • the master speaker sends the volume increase control instruction (i.e. the second control play instruction) to the sub-speaker in the bedroom to control the playing sound of the sub-speaker to change from small to large.
  • the audio data sent by the master speaker based on the spatial identifier is acquired.
  • the audio is played based on different spatial identifiers.
  • the master speaker can also control the change of volume of the sub-speaker playing audio according to the detection result of the sub-speaker in the subspace, so that the sub-speaker can achieve a gradual effect in volume and bring a good listening experience to the user.
  • Fig. 8 is a flow chart of playing audio based on audio data according to an illustrative embodiment. As shown in Fig. 8 , playing audio based on the audio data includes the following steps.
  • step S81 a system clock between the sub-speaker and the master speaker is calibrated.
  • Fig. 9 shows a schematic diagram of calibrating the system clock between the master speaker and the sub-speaker.
  • the master speaker and the sub-speaker also need to have an interaction of clock information.
  • the sub-speaker needs to calculate the system clock difference between itself and the master speaker, and adjust itself to the system clock consistent with the master speaker.
  • the sub-speaker sends time information to the master speaker through WiFi at time TB0, the master speaker receives the time information of the sub-speaker at time TA0 according to its own clock, then the master speaker sends time information to the sub-speaker through WiFi at time TA1, and the sub-speaker receives the time information of the master speaker at time TB1.
  • the sub-speaker compensates its own system clock with the clock difference ⁇ , and then it is calibrated to the system clock consistent with the master speaker. On the basis that all the speakers are aligned with the master speaker in clock, the audio frames at time T0 are played at time T0+1s, so that the synchronous play of all the speakers can be achieved.
  • step S82 based on the calibrated system clock and the audio data, the audio is played synchronously with the master speaker.
  • the system clock between the sub-speaker and the master speaker is calibrated. Based on the calibrated system clock and the audio data, the audio is played synchronously with the master speaker.
  • the synchronous audio play of the respective speakers is realized, and a good listening experience is brought to the user.
  • Fig. 10 shows a schematic diagram of a speaker play control.
  • the space where each speaker is located is divided, and the division of the subspace can be realized by user configuration and automatic perception.
  • the user configuration means that each time the user configures the network for the speaker, the user chooses the subspace in which the speaker is located. For example, the living room is chosen for speakers A and B, and the master bedroom is chosen for speakers C and D, so that speakers A and B mutually know that they are in a same subspace, and speakers C and D mutually know that they are in a same subspace.
  • the automatic perception means that the user does not need to choose, and the speakers perceive each other through ultrasonic communication.
  • speaker B only speaker B can receive the ultrasonic information sent by speaker A, and speaker C and speaker D cannot receive it due to wall blocking, so that speaker B knows that it is in the same subspace with speaker A.
  • Speaker A is used as the master speaker, the information of the spaces where all the devices are located is finally summarized into speaker A, and the IP addresses or other identification IDs of the devices in the respective subspaces are stored in the list. In this case, the relationship between the speakers is established.
  • speaker M detects the user by ultrasonic technology. If speaker M detects that the user is in subspace 1, it sends an instruction to the master speaker to inform that there is the user in subspace 1.
  • the master speaker When the master speaker receives the instruction from speaker M, it continues to send the audio data to speaker M and speaker N in subspace 1. If speaker M does not detect that the user is in subspace 1, it sends an instruction to the master speaker to inform that there is no user in subspace 1. When the master speaker receives the instruction from speaker M, it stops sending the audio data to speaker M and speaker N in subspace 1, and closes the channels between the master speaker and speakers M and N. Through the present disclosure, the play function of the speaker in the room without the user is stopped. Thus, on the one hand, the unnecessary waste of electric energy is saved, and on the other hand, it saves the occupation of network resources, reduces the network load and helps to improve the stability of the system.
  • the embodiments of the present disclosure provide the speaker play control method, which is applied to the sub-speakers in the respective divided subspaces, and the sub-speakers communicate with the master speaker.
  • the sub-speaker detects that there is the user in the space through ultrasonic technology, it sends the instruction that there is the user to the master speaker.
  • the sub-speaker detects that there is no user in the space through ultrasonic technology, it sends the instruction that there is no user to the master speaker.
  • the sub-speaker receives the audio data and the play control instruction from the master speaker, decodes the data and plays the decoded data, under the condition of informing the master speaker that there is the user in the space.
  • the embodiments of the present disclosure also provide a speaker play control device.
  • the speaker play control device provided by the embodiments of the present disclosure includes corresponding hardware structures and/or software modules for executing various functions.
  • the embodiments of the present disclosure can be realized in the form of hardware or a combination of hardware and computer software. Whether a function is executed by hardware or in a manner of computer software driving hardware depends on the specific application and design constraints of the technical solution. Those skilled in the art can use different methods to realize the described functions for each specific application, but this realization should not be considered beyond the scope of the technical solution of the embodiment of the present disclosure.
  • Fig. 11 is a block diagram of a speaker play control device according to an illustrative embodiment.
  • the device 100 can be provided as the master speaker according to the above embodiments, and include a determining unit 101 and a playing unit 102.
  • the determining unit 101 is configured to acquire audio data from a server in response to enabling an audio play function, and determine a target subspace in subspaces to which a plurality of sub-speakers belong.
  • the target subspace is a subspace where a user is currently located.
  • the playing unit 102 is configured to send the audio data to the sub-speaker in the target subspace and control the sub-speaker in the target subspace to play audio based on the audio data.
  • the determining unit 101 determines the target subspace in the subspaces to which the plurality of sub-speakers belong in the following ways: determining a user detection result, where the user detection result is determined by a set sub-speakers in the subspace based on detection of a human body activity of the user, and the user detection result includes a presence or an absence of the user; and determining the subspace to which the target sub-speaker whose detection result is the presence of the user belongs as the target subspace.
  • the plurality of sub-speakers correspond to spatial identifiers
  • the spatial identifier is used for identifying the subspace to which the sub-speaker belongs.
  • the playing unit 102 sends the audio data to the sub-speaker in the target subspace and control the sub-speaker in the target subspace to play audio based on the audio data in the following ways: determining a target spatial identifier corresponding to the target subspace; sending the audio data to the sub-speaker in the target subspace identified by the target spatial identifier, and controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data.
  • the playing unit 102 controls the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data in the following manners: sending a first control play instruction to the sub-speaker in the first target subspace, and controlling the sub-speaker in the first target subspace based on the first control play instruction to play audio based on the first control play instruction and the audio data, where the first control play instruction is used to control a playing sound of the sub-speaker to change from large to small; sending a second control play instruction to the sub-speaker in the second target subspace, and controlling the sub-speaker in the second target subspace based on the second control play instruction to play audio based on the second control play instruction and the audio data, where the second control play instruction is used to control the playing sound of the sub-speaker to change from small to large.
  • the playing unit 102 is also used to: determine a non-target subspace, which is a subspace without the user to which each set sub-speaker belongs; and stop sending the audio data to the non-target subspace.
  • Fig. 12 is a block diagram of a speaker play control device according to an illustrative embodiment.
  • the device 200 can be provided as the sub-speaker according to the above embodiments, and include an acquiring unit 201 and a playing unit 202.
  • the acquiring unit 201 is configured to acquire audio data sent by a master speaker.
  • the playing unit 202 is used to play audio based on the audio data.
  • the sub-speaker is a set sub-speaker, and the set sub-speaker is used to detect whether there is a user in the subspace to which a target sub-speaker belongs.
  • the playing unit 202 is also used to: detect a human body activity and determine a user detection result based on a human body activity detection result, where the user detection result includes a presence or an absence of a user; and send the user detection result to the master speaker.
  • the sub-speaker corresponds to a spatial identifier
  • the spatial identifier is used to identify the subspace to which the sub-speaker belongs.
  • the acquiring unit 201 acquires the audio data sent by the master speaker in a following way: acquiring the audio data sent by the master speaker based on the spatial identifier.
  • the playing unit 202 plays audio based on the audio data in the following ways: in response to that the spatial identifier is a first target spatial identifier, acquiring a first control play instruction sent by the master speaker, where a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and the first control play instruction is used to control a playing sound of the sub-speaker to change from large to small; based on the first control play instruction, playing audio based on the audio data in a manner that the playing sound changes from large to small.
  • the playing unit 202 plays audio based on the audio data in the following ways: in response to that the spatial identifier is a second target spatial identifier, acquiring a second control play instruction sent by the master speaker, where the second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves, and the second control instruction is used to control the playing sound of the sub-speaker to change from small to large; and based on the second control play instruction, playing audio based on the audio data in a manner that the playing sound changes from small to large.
  • the playing unit 202 plays audio based on the audio data in the following ways: calibrating a system clock between the sub-speaker and the master speaker; and based on the calibrated system clock and the audio data, playing audio synchronously with the master speaker.
  • Fig. 13 is a block diagram of a device for speaker play control according to an illustrative embodiment.
  • a device 300 may be a mobile phone, a computer, a digital broadcasting terminal, a messaging device, a game console, a tablet device, a medical device, a fitness device, a personal digital assistant and the like.
  • the device 300 may include one or more of the following components: a processing component 302, a memory 304, a power component 306, a multimedia component 308, an audio component 310, an input/output (I/O) interface 312, a sensor component 314, and a communication component 316.
  • the processing component 302 generally controls the overall operation of the device 300, such as operations associated with display, telephone call, data communication, camera operation and recording operation.
  • the processing component 302 may include one or more processors 320 to execute instructions to complete all or part of the steps of the method described above.
  • the processing component 302 can include one or more modules to facilitate the interaction between the processing component 302 and other components.
  • the processing component 302 can include a multimedia module to facilitate interaction between the multimedia component 308 and the processing component 302.
  • the memory 304 is configured to store various types of data to support operations in the device 300. Examples of the data include instructions for any application or method operating on the device 300, contact data, phone book data, messages, pictures, videos, and the like.
  • the memory 304 can be realized by any type of volatile or nonvolatile memory device or their combination, such as static random access memory (SRAM), electrically erasable programmable read-only memory (EEPROM), erasable programmable read-only memory (EPROM), programmable read-only memory (PROM), read-only memory (ROM), magnetic memory, flash memory, magnetic disk or optical disk.
  • SRAM static random access memory
  • EEPROM electrically erasable programmable read-only memory
  • EPROM erasable programmable read-only memory
  • PROM programmable read-only memory
  • ROM read-only memory
  • magnetic memory flash memory
  • flash memory magnetic disk or optical disk.
  • the power component 306 provides power for various components of the device 300.
  • the power component 306 may include a power management system, one or more power sources, and other components associated with generating, managing and distributing power for the device 300.
  • the multimedia component 308 includes a screen that provides an output interface between the device 300 and the user.
  • the screen may include a liquid crystal display (LCD) and a touch panel (TP). If the screen includes a touch panel, the screen may be implemented as a touch screen to receive an input signal from the user.
  • the touch panel includes one or more touch sensors to sense touch, sliding and gestures on the touch panel. The touch sensor may not only sense the boundary of a touch or sliding action, but also detect the duration and pressure related to the touch or sliding operation.
  • the multimedia component 308 includes a front camera and/or a rear camera. When the device 300 is in an operation mode, such as a shooting mode or a video mode, the front camera and/or the rear camera can receive external multimedia data. Each front camera and rear camera can be a fixed optical lens system or have a focal length and an optical zoom capability.
  • the audio component 310 is configured to output and/or input audio signals.
  • the audio component 310 includes a microphone (MIC) configured to receive external audio signals when the device 300 is in an operation mode, such as a call mode, a recording mode and a voice recognition mode.
  • the received audio signal may be further stored in the memory 304 or transmitted via the communication component 316.
  • the audio component 310 further includes a speaker for outputting audio signals.
  • the I/O interface 312 provides an interface between the processing component 302 and peripheral interface modules, which can be keyboards, click wheels, buttons, etc. These buttons may include, but are not limited to, a home button, a volume button, a start button and a lock button.
  • the sensor assembly 314 includes one or more sensors for providing various aspects of the state evaluation for the device 300.
  • the sensor component 314 can detect the on/off state of the device 300, the relative positioning of components, such as the display and keypad of the device 300, the position change of the device 300 or a component of the device 300, the presence or absence of user contact with the device 300, the orientation or acceleration/deceleration of the device 300 and the temperature change of the device 300.
  • the sensor assembly 314 may include a proximity sensor configured to detect the presence of a nearby object without any physical contact.
  • the sensor assembly 314 may also include an optical sensor, such as a CMOS or CCD image sensor, for use in imaging applications.
  • the sensor assembly 314 may further include an acceleration sensor, a gyro sensor, a magnetic sensor, a pressure sensor or a temperature sensor.
  • the communication component 316 is configured to facilitate wired or wireless communication between the device 300 and other devices.
  • the device 300 can access a wireless network based on communication standards, such as WiFi, 2G or 3G, or a combination thereof.
  • the communication component 316 receives a broadcast signal or broadcast related information from an external broadcast management system via a broadcast channel.
  • the communication component 316 further includes a near field communication (NFC) module to facilitate short-range communication.
  • NFC near field communication
  • the NFC module can be implemented based on radio frequency identification (RFID) technology, infrared data association (IrDA) technology, ultra-wideband (UWB) technology, Bluetooth (BT) technology and other technologies.
  • RFID radio frequency identification
  • IrDA infrared data association
  • UWB ultra-wideband
  • Bluetooth Bluetooth
  • the device 300 may be implemented by one or more application specific integrated circuits (ASIC), digital signal processors (DSP), digital signal processing devices (DSPD), programmable logic devices (PLD), field programmable gate arrays (FPGA), controllers, microcontrollers, microprocessors or other electronic components, for performing the above methods.
  • ASIC application specific integrated circuits
  • DSP digital signal processors
  • DSPD digital signal processing devices
  • PLD programmable logic devices
  • FPGA field programmable gate arrays
  • controllers microcontrollers, microprocessors or other electronic components, for performing the above methods.
  • non-transitory computer-readable storage medium including instructions, such as the memory 304 including instructions, and the instructions can be executed by the processor 320 of the device 300 to complete the above method.
  • the non-transitory computer-readable storage medium can be ROM, random access memory (RAM), CD-ROM, magnetic tape, floppy disk, optical data storage device, etc.
  • first and second are used to describe various information, but the information should not be limited to these terms. These terms are only used to distinguish the same type of information from each other and do not indicate a specific order or importance. In fact, the expressions “first” and “second” can be used interchangeably.
  • the first information may also be called the second information, and similarly, the second information may also be called the first information.
  • connection includes direct connection between two without other components, and indirect connection between two with other components.

Landscapes

  • Physics & Mathematics (AREA)
  • Engineering & Computer Science (AREA)
  • Acoustics & Sound (AREA)
  • Signal Processing (AREA)
  • Health & Medical Sciences (AREA)
  • General Health & Medical Sciences (AREA)
  • Otolaryngology (AREA)
  • Circuit For Audible Band Transducer (AREA)

Abstract

The present disclosure relates to a sound box playing control method, a sound box playing control apparatus, and a storage medium. The sound box playing control method comprises: in response to an audio playing function being enabled, acquiring audio data from a serving end, and determining a target subspace from among subspaces to which a plurality of sub-sound boxes belong, wherein the target subspace is a subspace where a user is currently located; and sending the audio data to a sub-sound box in the target sub-space, and controlling the sub-sound box in the target sub-space to perform audio playing on the basis of the audio data. By means of the present disclosure, the waste of energy is reduced; moreover, channel resources between a main control sound box and a sub-sound box, which does not play audio data, are released, thereby providing a more stable auditory experience for users.

Description

    FIELD
  • The present disclosure relates to the field of speakers, in particular to a speaker play control method, a speaker play control device and a storage medium.
  • BACKGROUND
  • With the development of smart home technology and people's pursuit of high-quality home life, the whole house speaker brings shocking auditory enjoyment to people's home life. A plurality of smart speakers are connected to each other through WiFi and play music synchronously, so that users can have an immersive listening experience while walking. The terminal speaker gets audio data from the server and transmits it to the speakers in all rooms through the wireless network. Through a certain synchronization mechanism, the same audio frame can be played at the same time.
  • However, the users usually only play music in one or part of the rooms, and playing music in the rooms without the users will lead to the loss of electric energy on the one hand. On the other hand, since the speakers sending the data will occupy the network channel, the more network resources are occupied, the more likely it will lead to instability of the system. For example, the problem such as unsynchronized play will occur.
  • SUMMARY
  • In order to overcome the problems existing in the related art, the present disclosure provides a speaker play control method, a speaker play control device and a storage medium.
  • According to embodiments of a first aspect of the present disclosure, there is provided a speaker play control method, which is applied to a master speaker, and the master speaker is configured to control a plurality of sub-speakers, and the plurality of sub-speakers belong to different subspaces of a same space. The method includes: in response to enabling an audio play function, acquiring audio data from a server, and determining a target subspace in the subspaces to which the plurality of sub-speakers belong, wherein the target subspace is a subspace where a user is currently located; and sending the audio data to a sub-speaker in the target subspace, and controlling the sub-speaker in the target subspace to play audio based on the audio data.
  • In an embodiment, the determining the target subspace in the subspaces to which the plurality of sub-speakers belong includes: determining a user detection result, wherein the user detection result is determined by a set sub-speaker in the subspace based on detection of a human body activity of the user, and the user detection result comprises a presence or an absence of the user; and determining a subspace to which a target sub-speaker whose detection result is the presence of the user belongs as the target subspace.
  • In an embodiment, the plurality of sub-speakers correspond to spatial identifiers, and the spatial identifier is configured to identify the subspace to which the sub-speaker belongs; the sending the audio data to the sub-speaker in the target subspace and controlling the sub-speaker in the target subspace to play audio based on the audio data includes: determining a target spatial identifier corresponding to the target subspace; and sending the audio data to the sub-speaker in the target subspace identified by the target spatial identifier, and controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data.
  • In an embodiment, the target spatial identifier comprises a first target spatial identifier and a second target spatial identifier, a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves; the controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data includes: sending a first control play instruction to the sub-speaker in the first target subspace, and controlling the sub-speaker in the first target subspace based on the first control play instruction to play audio based on the first control play instruction and the audio data, wherein the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and sending a second control play instruction to the sub-speaker in the second target subspace, and controlling the sub-speaker in the second target subspace based on the second control play instruction to play audio based on the second control play instruction and the audio data, wherein the second control play instruction is configured to control the playing sound of the sub-speaker to change from small to large.
  • In an embodiment, the method further includes: determining a non-target subspace, wherein the non-target subspace is a subspace without the user to which each set sub-speaker belongs; and stopping sending the audio data to the non-target subspace.
  • According to embodiments of a second aspect of the present disclosure, there is provided a speaker play control method, which is applied to a sub-speaker, and a subspace to which the sub-speaker belongs is a target subspace, and the target subspace is a subspace where a user is currently located. The method includes: acquiring audio data sent by a master speaker; and playing audio based on the audio data.
  • In an embodiment, the sub-speaker is a set sub-speaker, and the set sub-speaker is configured to detect whether there is the user in a subspace to which a target sub-speaker belongs, and the method further includes: detecting a human body activity, and determining a user detection result based on a human body activity detection result, wherein the user detection result comprises a presence or an absence of the user; and sending the user detection result to the master speaker.
  • In an embodiment, the sub-speaker corresponds to a spatial identifier, and the spatial identifier is configured to identify the subspace to which the sub-speaker belongs; and the acquiring the audio data sent by the master speaker includes: acquiring the audio data sent by the master speaker based on the spatial identifier.
  • In an embodiment, the playing audio based on the audio data includes: in response to that the spatial identifier is a first target spatial identifier, acquiring a first control play instruction sent by the master speaker, wherein a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and based on the first control play instruction, playing audio based on the audio data in a manner that the playing sound changes from large to small.
  • In an embodiment, the playing audio based on the audio data includes: in response to that the spatial identifier is a second target spatial identifier, acquiring a second control play instruction sent by the master speaker, wherein a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves, and the second control play instruction is configured to control a playing sound of the sub-speaker to change from small to large; and based on the second control play instruction, playing audio based on the audio data in a manner that the playing sound changes from small to large.
  • In an embodiment, the playing audio based on the audio data includes: calibrating a system clock between the sub-speaker and the master speaker; and based on the calibrated system clock and the audio data, playing audio synchronously with the master speaker.
  • According to embodiments of a third aspect of the present disclosure, there is provided a speaker play control device, which includes: a determining unit configured to acquire audio data from a server in response to enabling an audio play function, and determining a target subspace in subspaces to which a plurality of sub-speakers belong, wherein the target subspace is a subspace where a user is currently located; and a playing unit configured to send the audio data to a sub-speaker in the target subspace and control the sub-speaker in the target subspace to play audio based on the audio data.
  • In an embodiment, the determining unit determines the target subspace in the subspaces to which the plurality of sub-speakers belong in a following manner: determining a user detection result, wherein the user detection result is determined by a set sub-speaker in the subspace based on detection of a human body activity of the user, and the user detection result comprises a presence or an absence of the user; and determining a subspace to which a target sub-speaker whose detection result is the presence of the user belongs as the target subspace.
  • In an embodiment, the plurality of sub-speakers correspond to spatial identifiers, and the spatial identifier is configured to identify the subspace to which the sub-speaker belongs; the playing unit sends the audio data to the sub-speaker in the target subspace and controls the sub-speaker in the target subspace to play audio based on the audio data in a following manner: determining a target spatial identifier corresponding to the target subspace; and sending the audio data to the sub-speaker in the target subspace identified by the target spatial identifier, and controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data.
  • In an embodiment, the target spatial identifier comprises a first target spatial identifier and a second target spatial identifier, wherein a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves; the playing unit controls the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data in a following manner; sending a first control play instruction to the sub-speaker in the first target subspace, and controlling the sub-speaker in the first target subspace based on the first control play instruction to play audio based on the first control play instruction and the audio data, wherein the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and sending a second control play instruction to the sub-speaker in the second target subspace, and controlling the sub-speaker in the second target subspace based on the second control play instruction to play audio based on the second control play instruction and the audio data, wherein the second control play instruction is configured to control the playing sound of the sub-speaker to change from small to large.
  • In an embodiment, the playing unit is further configured to: determine a non-target subspace, wherein the non-target subspace is a subspace without the user to which each set sub-speaker belongs; and stop sending the audio data to the non-target subspace.
  • According to embodiments of a fourth aspect of that present disclosure, there is provided a speaker play control device, which includes: an acquiring unit configured to acquire audio data sent by a master speaker; and a playing unit configured to play audio based on the audio data.
  • In an embodiment, the sub-speaker is a set sub-speaker, and the set sub-speaker is configured to detect whether there is the user in a subspace to which a target sub-speaker belongs, and the playing unit is further configured to: detect a human body activity, and determining a user detection result based on a human body activity detection result, wherein the user detection result comprises a presence or an absence of the user; and send the user detection result to the master speaker.
  • In an embodiment, the sub-speaker corresponds to a spatial identifier, and the spatial identifier is configured to identify the subspace to which the sub-speaker belongs; the acquiring unit acquires the audio data sent by the master speaker in a following way: acquiring the audio data sent by the master speaker based on the spatial identifier.
  • In an embodiment, the playing unit plays audio based on the audio data in a following manner: in response to that the spatial identifier is a first target spatial identifier, acquiring a first control play instruction sent by the master speaker, wherein a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and based on the first control play instruction, playing audio based on the audio data in a manner that the playing sound changes from large to small.
  • In an embodiment, the playing unit plays audio based on the audio data in a following manner: in response to that the spatial identifier is a second target spatial identifier, acquiring a second control play instruction sent by the master speaker, wherein a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves, and the second control play instruction is configured to control a playing sound of the sub-speaker to change from small to large; and based on the second control play instruction, playing audio based on the audio data in a manner that the playing sound changes from small to large.
  • In an embodiment, the playing unit plays audio based on the audio data in a following manner: calibrating a system clock between the sub-speaker and the master speaker; and based on the calibrated system clock and the audio data, playing audio synchronously with the master speaker.
  • According to embodiments of a fifth aspect of that present disclosure, there is provided a speaker play control device, which includes: a processor; and a memory for storing instructions executable by the processor. The processor is configured to perform the method described in the first aspect or any one of the embodiments of the first aspect.
  • According to embodiments of a sixth aspect of the present disclosure, there is provided a speaker play control device, which includes: a processor; and a memory for storing instructions executable the processor. The processor is configured to perform the method described in the second aspect or any one of the embodiments of the second aspect.
  • According to embodiments of a seventh aspect of the present disclosure, there is provided a computer-readable storage medium in which instructions are stored. When the instructions in the storage medium are executed by a processor of a network device, the network device is enabled to perform the method described in the first aspect or any one of the embodiments of the first aspect.
  • According to embodiments of an eighth aspect of the present disclosure, there is provided a computer-readable storage medium in which instructions are stored. When the instructions in the storage medium are executed by a processor of a terminal, the terminal is enabled to perform the method described in the second aspect or any one of the embodiments of the second aspect.
  • The technical solution provided by the embodiments of the present disclosure can include the following beneficial effects. When it is determined that the speaker is enabled with the audio play function, the audio data is acquired from the server through the master speaker, and the space where the user is currently located is determined. Further, the master speaker sends the audio data acquired from the server to the sub-speaker in the space where the user is currently located, and the sub-speaker plays audio based on the audio data. Based on this, the speaker in the space selectively plays the audio data according to the space where the user is currently located, thus reducing the waste of energy. That is, in the current space, the speaker which does not detect that the user is in this space is in a non-working state. At the same time, the channel resources between the master speaker and the sub-speakers that do not play the audio data are released, thus providing the user with a more stable listening experience.
  • It is to be understood that both the foregoing general description and the following detailed description are illustrative and explanatory only, and are not restrictive of the present disclosure.
  • BRIEF DESCRIPTION OF THE DRAWINGS
  • The accompanying drawings, which are incorporated in and constitute a part of the specification, illustrate embodiments consistent with the present disclosure and together with the description, serve to explain the principles of the present disclosure.
    • Fig. 1 is a flow chart of a speaker play control method according to an illustrative embodiment.
    • Fig. 2 is a flow chart of determining a target subspace according to an illustrative embodiment.
    • Fig. 3 is a flow chart of controlling a sub-speaker in a target subspace to play audio based on audio data according to an illustrative embodiment.
    • Fig. 4 is a flow chart of controlling a sub-speaker in a target subspace to play audio based on audio data according to an illustrative embodiment.
    • Fig. 5 is a flow chart of a speaker play control method according to an illustrative embodiment.
    • Fig. 6 is a flow chart of detecting whether a user exists in a subspace to which a target sub-speaker belongs according to an illustrative embodiment.
    • Fig. 7 is a flow chart of acquiring audio data sent by a master speaker according to an illustrative embodiment.
    • Fig. 8 is a flow chart of playing audio based on audio data according to an illustrative embodiment.
    • Fig. 9 shows a schematic diagram of calibrating a system clock between a master speaker and a sub-speaker.
    • Fig. 10 shows a schematic diagram of speaker play control.
    • Fig. 11 is a block diagram of a speaker play control device according to an illustrative embodiment.
    • Fig. 12 is a block diagram of a speaker play control device according to an illustrative embodiment.
    • Fig. 13 is a block diagram of a device for speaker play control according to an illustrative embodiment.
    DETAILED DESCRIPTION
  • Reference will now be made in detail to illustrative embodiments, examples of which are illustrated in the accompanying drawings. When the following description refers to the drawings, unless otherwise indicated, the same numbers in different drawings indicate the same or similar elements. The implementations described in the following illustrative embodiments do not represent all the implementations consistent with the present disclosure.
  • In the accompanying drawings, the same or similar reference numerals indicate the same or similar elements or elements with the same or similar functions throughout the specification. The described embodiments are part of the embodiments of the present disclosure, but not all the embodiments. The embodiments described below with reference to the accompanying drawings are illustrative and are intended to explain the present disclosure, and should not be construed as limiting the present disclosure. Based on the embodiments in the present disclosure, all other embodiments acquired by those ordinary skilled in the art without creative work belong to the scope of protection of the present disclosure. Hereinafter, the embodiments of the present disclosure will be described in detail with reference to the accompanying drawings.
  • A speaker play control method provided by embodiments of the present disclosure may be applied to a plurality of smart devices connected by wireless communication technology, so as to realize an application scene of sharing music between users. In particular, the plurality of smart devices are in a set space. The plurality of smart devices belong to different subspaces respectively, and the plurality of smart devices include a master smart device and a sub-smart device. The smart device may be a smart speaker or other smart devices.
  • With the popularity of the smart speaker and the increasing demand of the user listening to music, major companies are committed to improving the music play function of their speakers. One of the breakthrough progresses is that the speakers are interconnected through a wireless network to realize the whole house play function. A plurality of smart speakers are connected to each other through WiFi and play music synchronously, so that the user can have an immersive listening experience while walking. The speakers get the music data from the server and send it to the speakers in all rooms through the wireless network. Through a certain synchronization mechanism, the same audio frame can be played at the same time, so that all the smart speakers in the family can play music synchronously.
  • In the related art, the information interaction among the plurality of smart speakers is realized through WiFi interconnection. The master speaker gets the audio data from the server, frames the audio data by an equal time length, marks corresponding time information for each audio frame, and packages it and sends it to other sub-speakers in all rooms through WiFi. The other sub-speakers receive the audio frame packets sent by the master speaker through WiFi and analyze the audio data and the time information. At the same time of interaction of the audio frame packet, the other sub-speakers should have a system clock interaction with the master speaker, so as to achieve the alignment of the master speaker and the other sub-speakers in terms of the system clock, thus ensuring that the data processing of the master speaker and the other sub-speakers is on the same clock reference. Based on the clock synchronization of the master speaker and the other sub-speakers, the audio data is processed respectively. The master speaker and the other sub-speakers are set to play the same audio frame after a fixed time, so that the plurality of speakers can play the audio synchronously, thus achieving the listening experience of the whole house play. However, in the family, the user usually only plays music in one or part of the rooms, and playing music in the room without the user will waste electric energy on the one hand. On the other hand, sending the data will occupy the network channel. The more network resources are occupied, the more likely it will lead to instability of the system. For example, the plurality of speakers cannot play the audio data synchronously.
  • In view of this, the present disclosure provides a speaker play control method. When it is determined that the speaker is enabled with the audio play function, the audio data is acquired from the server through the master speaker, and a target subspace is determined in subspaces to which a plurality of sub-speakers belong. The target subspace is a subspace where the user is currently located, and the subspace is one of a plurality of spaces into which a complete space is divided. Further, the master speaker sends the audio data acquired from the server to the sub-speaker in the space where the user is currently located, and the sub-speaker plays audio based on the audio data. Based on this, the speaker in the space selectively plays the audio data according to the space where the user is currently located, thus reducing the waste of energy, that is, for a current space, when no user is detected in this space, the speaker in this space is in a non-working state. At the same time, the channel resources between the master speaker and the sub-speakers that do not play the audio data are released, which reduces the network load as much as possible, thus improving the stability of the system and providing the user with a more stable listening experience. Therefore, compared with the way of controlling the speaker to play in the related art, the speaker play control method provided by the present disclosure is more flexible and intelligent.
  • Fig. 1 is a flow chart of a speaker play control method according to an illustrative embodiment. As shown in Fig. 1, the method is applied to a master speaker and includes the following steps.
  • In the following disclosed embodiments, a speaker communicating with the server and other speakers is called a master speaker, and a speaker communicating with the master speaker is called a sub-speaker. The master speaker is used to control a plurality of sub-speakers, and different sub-speakers are included in different subspaces of the same space. One or more different sub-speakers may be included in the same subspace.
  • In step S11, when it is determined that the master speaker is enabled with an audio play function, audio data is acquired from a server, and a target subspace is determined in subspaces to which a plurality of sub-speakers belong.
  • The target subspace is a subspace where the user is currently located. The subspace may be understood as any room, such as a living room, a kitchen, a bedroom and so on.
  • In embodiments of the present disclosure, when it is determined that the user has enabled the audio play functions of all the speakers in the rooms, the master speaker acquires the audio data from the server through a wireless transmission technology, and determines the room where the user is currently located, by receiving a signal sent by the sub-speaker in the subspace.
  • In step S12, the audio data is sent to the sub-speaker in the target subspace, and the sub-speaker in the target subspace is controlled to play audio based on the audio data.
  • In embodiments of the present disclosure, in case that the master speaker determines which space has the user, it sends the audio data acquired from the server to the sub-speaker in the space where the user exists, and controls the sub-speaker in the space to play the decoded audio data.
  • In the present disclosure, the master speaker determines a non-target subspace through an instruction sent by the sub-speaker in the subspace, and the non-target subspace is a subspace to which a set sub-speaker whose detection result is the absence of the user belongs. In case that the master speaker determines the non-target subspace, it stops sending the audio data to the sub-speaker in the non-target subspace.
  • In embodiments of the present disclosure, in case that the master speaker determines that there is no user in the subspace, the master speaker stops sending the audio data to the sub-speaker in the space and closes a channel between the master speaker and the sub-speaker in the space. The channel resources between the master speaker and the sub-speaker that does not play the audio data are released, which improves the stability of the system and provides the user with a more stable listening experience.
  • In the present disclosure, after it is determined that all the speakers are enabled with the audio play function, the audio data is acquired from the server, and the target subspace is determined in the subspaces to which the plurality of sub-speakers belong. The audio data is sent to the sub-speaker in the target subspace, and the sub-speaker in the target subspace is controlled to play audio based on the audio data. Through the present disclosure, the speaker in the space selectively plays the audio data according to the space where the user is currently located, thereby reducing the waste of energy. It can be understood that only when there is the user in the space, the master speaker sends the audio data to the sub-speaker in the space.
  • Based on the above embodiments, it can be seen that it is a key step that the master speaker determines the target subspace. Therefore, in the following disclosed embodiments, how the master speaker determines the target subspace will be specifically explained.
  • Fig. 2 is a flow chart of determining a target subspace according to an illustrative embodiment. As shown in fig. 2, determining the target subspace in the subspaces to which the plurality of sub-speakers belong includes the following steps.
  • In step S21, a user detection result is determined.
  • The user detection result is determined by a set sub-speaker in the subspace based on the detection of the human body activity of the user, and the user detection result includes the presence or absence of the user.
  • In step S22, the subspace to which a target sub-speaker whose detection result is the presence of the user belongs is determined as the target subspace.
  • In embodiments of the present disclosure, if the sub-speaker set in the subspace detects the human body activity of the user in the space, the master speaker receives an instruction confirming the presence of the user from the sub-speaker, and determines the subspace with the user to which the target sub-speaker belongs as the target subspace. If the sub-speaker set in the subspace does not detect the human body activity of the user in the space, the master speaker receives an instruction for confirming the absence of the user from the sub-speaker.
  • In the present disclosure, the user detection result is determined, and the subspace to which the target sub-speaker whose detection result is the presence of the user belongs is determined as the target subspace. Through the present disclosure, based on the detection result instruction transmitted from the sub-speaker to the master speaker, the master speaker sends a corresponding instruction to control the sub-speaker, thus realizing the interaction between the sub-speaker and the master speaker.
  • Fig. 3 is a flow chart of controlling a sub-speaker in a target subspace to play audio based on audio data according to an illustrative embodiment. As shown in Fig. 3, sending the audio data to the sub-speaker in the target subspace and controlling the sub-speaker in the target subspace to play audio based on the audio data includes the following steps.
  • In step S31, a target spatial identifier corresponding to the target subspace is determined.
  • The plurality of sub-speakers correspond to spatial identifiers, and the spatial identifier is used to identify the subspace to which the sub-speaker belongs.
  • In embodiments of the present disclosure, the room environment is divided into a plurality of subspaces, such as a kitchen, a living room, a bedroom and other subspaces. When configuring the network for the first time, the user can select the space where the speaker is located. For example, the user can select the space where the speaker is currently located through a resizable list on a display screen of the speaker, so that each speaker knows the space environment where it is currently located. For example, speaker A and speaker B are currently placed in the living room, and speaker C and speaker D are currently placed in the master bedroom and so on. Secondly, the sub-speakers in the target subspace perceive each other through ultrasonic communication to determine the subspace to which each sub-speaker belongs. The sub-speaker sends its own spatial identifier to the master speaker, and the master speaker stores IP addresses or other identification IDs of all the sub-speakers in the subspace in a list of devices, so that a target spatial identifier corresponding to the target subspace can be determined.
  • In step S32, the audio data is sent to the sub-speaker in the target subspace identified by the target spatial identifier, and the sub-speaker in the target subspace are controlled based on the target spatial identifier to play audio based on the audio data.
  • In an example, the user can select one of the sub-speakers in the target subspace in advance, and this sub-speaker detects whether there is the user activity in the space, and sends the detection result to the master speaker. It is also possible to set the speaker with the best performance to detect whether there is the user activity in the space through the mutual perception of the sub-speakers in the target subspace, and send the detection result to the set speaker of the master speaker. Assuming that the user is currently active in the master bedroom and the sub-speaker set in the master bedroom detects the human body activity, the speaker sends a confirmation instruction to the master speaker, and the master speaker sends the audio data to the sub-speaker in the master bedroom and controls the sub-speaker in the master bedroom to play audio.
  • In the present disclosure, the target spatial identifier corresponding to the target subspace is determined. The audio data is sent to the sub-speaker in the target subspace identified by the target spatial identifier, and the sub-speaker in the target subspace is controlled based on the target spatial identifier to play audio based on the audio data. Through the present disclosure, the master speaker can accurately control the sub-speaker in the space where the user exists to play audio.
  • Fig. 4 is a flow chart of controlling a sub-speaker in a target subspace to play audio based on audio data according to an illustrative embodiment. As shown in Fig. 4, controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data includes the following steps.
  • In step S41, a first control play instruction is sent to a sub-speaker in a first target subspace, and the sub-speaker in the first target subspace is controlled based on the first control play instruction to play audio based on the first control play instruction and the audio data.
  • The first control play instruction is used to control a playing sound of the sub-speaker to change from large to small.
  • In embodiments of the present disclosure, when the user leaves the living room (i.e., the first target subspace), the sub-speaker in the living room cannot detect the user's human body activity at this time, and the sub-speaker sends the instruction for confirming the absence of the user to the master speaker. After receiving this instruction, the master speaker sends a volume reduction control instruction (i.e. the first control play instruction) to the sub-speaker in the living room, and controls the playing sound of the sub-speaker to change from large to small until there is no sound.
  • In step S42, a second control play instruction is sent to the sub-speaker in a second target subspace, and the sub-speaker in the second target subspace is controlled to play audio based on the second control play instruction and the audio data.
  • The second control play instruction is used to control the playing sound of the sub-speaker to change from small to large.
  • In embodiments of the present disclosure, when the user walks into the bedroom (i.e., the second target subspace), the sub-speaker in the bedroom detects the human body activity of the user at this time, and the sub-speaker sends the instruction for confirming the presence of the user to the master speaker. After receiving this instruction, the master speaker sends a volume increase control instruction (i.e. the second control play instruction) to the sub-speaker in the bedroom to control the playing sound of the sub-speaker to change from small to large.
  • In the present disclosure, the first control play instruction is sent to the sub-speaker in the first target subspace, and the sub-speaker in the first target subspace is controlled based on the first control play instruction to play audio based on the first control play instruction and the audio data. The second control play instruction is sent to the sub-speaker in the second target subspace, and the sub-speaker in the second target subspace is controlled based on the second control play instruction to play audio based on the second control play instruction and the audio data. Through the present disclosure, the master speaker can also control the change of volume of the audio played by the sub-speaker according to the detection result of the sub-speaker in the subspace, so that the sub-speaker can achieve a gradual effect in volume and bring a good listening experience to the user.
  • The embodiments of the present disclosure provide the speaker play control method, which is applied to the master speaker, and the master speaker communicates with the server and receives the audio data from the server. The received audio data is decoded to obtain the decoded audio data. Then, the decoded audio data is divided according to a preset duration to obtain audio frame data. According to the received instructions of the sub-speakers in the respective subspaces, further operations are carried out. If the instruction that the sub-speaker informs the presence of the user is received, the audio data and the play control instruction are continued to be sent to the sub-speaker in the space, and the clock data for calibration also needs to be transmitted while the audio data is transmitted. If the instruction that the sub-speaker informs the absence of the user is received, the audio data is stopped from being sent to the sub-speaker in the space. According to the embodiments of the present disclosure, the speaker in the space selectively plays the audio data according to the space where the user is currently located, thereby reducing the waste of energy. At the same time, the channel resources between the master speaker and the sub-speakers that do not play the audio data are released, which improves the stability of the system.
  • Fig. 5 is a flow chart of a speaker play control method according to an illustrative embodiment, which is applied to a sub-speaker as shown in Fig. 5 and includes the following steps.
  • In step S51, the sub-speaker in the target subspace acquires the audio data sent by the master speaker.
  • In step S52, audio is played based on the audio data.
  • In the present disclosure, the audio data sent by the master speaker is acquired, and the audio is played based on the audio data. According to the present disclosure, the sub-speaker is controlled by the master speaker, and the interaction between the sub-speaker and the master speaker is realized.
  • Fig. 6 is a flow chart of detecting whether a user exists in a subspace to which a target sub-speaker belongs according to an illustrative embodiment. As shown in Fig. 6, the sub-speaker is a set sub-speaker, and the set sub-speaker is used to detect whether the user exists in the subspace to which the target sub-speaker belongs, and the method includes the following steps.
  • In step S61, the human body activity is detected, and the user detection result is determined based on a detection result of the human body activity.
  • The user detection results include the presence or absence of the user.
  • In embodiments of the present disclosure, one sub-speaker is selected randomly or selected according to the performance of the sub-speaker in each subspace to transmit and receive an ultrasonic wave. When the audio play functions of all the speakers are enabled, the selected sub-speaker in each subspace emits the ultrasonic wave to detect the human body activity of the user. In embodiments of the present disclosure, a differential channel impulse response (dCIR) is used to detect the human body activity of the user. The essence of this detection method is through calculating the dCIR in the environment in real-time. When the user is not active in subspace A at present, the dCIR value of the sub-speaker in subspace A receiving the ultrasonic wave back approaches zero. At this time, the sub-speaker set in the target subspace will send the instruction that the user does not exist in the space to the master speaker. After receiving the instruction, the master speaker sends the corresponding instruction operation to the sub-speaker. When the user is currently in subspace B, the dCIR value of the sub-speaker in subspace B receiving the ultrasonic wave back will reflect the amplitude change on the dCIR. At this time, if the dCIR amplitude of the sub-speaker set in the target subspace changes, the speaker will send the instruction that the user exists in the space to the master speaker. After receiving the instruction, the master speaker sends the corresponding instruction operation to the sub-speaker. By detecting the overall amplitude state of the dCIR, it can be detected whether there is the user in the current subspace.
  • In embodiments of the present disclosure, the application principle of the dCIR is based on the following formula: s 1 s 2 s L s 2 s 3 s L + 1 s P s P + 1 s P + L 1 TrainingMatrix h 1 h 2 h L CIR = r L + 1 r L + 2 r L + P R n .
    Figure imgb0001
  • In the above formula, S is an ultrasonic signal emitted by a loudspeaker, r is the ultrasonic signal collected by a microphone, and H is a CIR vector. The calculation formula of h is: h = (STS)-1 SR. The dCIR is described as: dCIR m = = h m - h m-1, where m is a current frame. The amplitude statistic of the dCIR is: AMP dCI R = i = 1 L h ^ i
    Figure imgb0002
    . When the AMPdCIR is greater than a set threshold, it is determined that there is the user in the current subspace, and the set threshold may be 3 after being verified by experiments.
  • In step S62, the user detection result is sent to the master speaker.
  • In embodiments of the present disclosure, when the sub-speaker in the subspace determines that the user is currently in the same room as the sub-speaker, the instruction is sent to the master speaker, informing the master speaker that it can continue to send the audio frame data to the sub-speaker in the subspace. In this case, when the master speaker receives the instruction sent from the sub-speaker, it can be determined that the user detection result determined by the set sub-speaker in this subspace is that there is the user. The master speaker determines the subspace to which the target speaker whose detection result is the presence of the user belongs as the target subspace.
  • In embodiments of the present disclosure, the sub-speaker in the subspace may also detect that there is no user in the current subspace.
  • In the present disclosure, non-target subspace information is sent to the master speaker, and a non-target subspace is a subspace without the user to which each set sub-speaker belongs. The master speaker is controlled to stop sending the audio data to the non-target subspace.
  • In embodiments of the present disclosure, when the dCIR value of the ultrasonic wave received back by the sub-speaker in the subspace approaches zero, it is determined that there is no user in the current subspace. When it is determined that there is no user in this subspace, the sub-speaker will send the instruction to the master speaker to inform that there is no user in the space to which it belongs.
  • In embodiments of the present disclosure, when the master speaker receives the non-target subspace information sent from the sub-speaker, it is determined that there is no user in the space where the sub-speaker is located. In order to avoid waste of resources, the master speaker stops sending the audio data to the non-target subspace.
  • In the present disclosure, the human body activity is detected, and the user detection result is determined based on the human body activity detection result. The user detection result is sent to the master speaker. Through the present disclosure, the audio data is stopped in time from being sent to the sub-speaker in the subspace where the user does not exist, and the channel resources are released dynamically, thus ensuring the stability of the system and providing a better listening experience.
  • Fig. 7 is a flow chart of acquiring audio data sent by a master speaker according to an illustrative embodiment. As shown in Fig. 7, the acquiring audio data sent by the master speaker includes the following steps.
  • In step S71, the audio data sent by the master speaker based on the spatial identifier is acquired.
  • In step S72, the audio is played based on different spatial identifiers.
  • In the present disclosure, in response to that the spatial identifier is a first target spatial identifier, the first control play instruction sent by the master speaker is acquired, and the first target subspace identified by the first target spatial identifier is the subspace where the user is before the user moves, and the first control play instruction is used to control the playing sound of the sub-speaker to change from large to small. Based on the first control play instruction, the audio is played based on the audio data in a manner that the playing sound changes from large to small.
  • In embodiments of the present disclosure, when the user leaves the living room (i.e., the first target subspace), the sub-speaker in the living room cannot detect the user's human body activity at this time, and the sub-speaker sends the instruction for confirming the absence of the user to the master speaker. After receiving the instruction, the master speaker sends the volume reduction control instruction (i.e. the first control play instruction) to the sub-speaker in the living room, and controls the playing sound of the sub-speaker to change from large to small until there is no sound.
  • In the present disclosure, in response to that the spatial identifier is a second target spatial identifier, the second control play instruction sent by the master speaker is acquired, and the second target subspace identified by the second target spatial identifier is the subspace where the user is after the user moves, and the second control play instruction is used to control the playing sound of the sub-speaker to change from small to large. Based on the second control play instruction, the audio is played based on the audio data in a manner that the playing sound changes from small to large.
  • In embodiments of the present disclosure, when the user walks into the bedroom (i.e., the second target subspace), the sub-speaker in the bedroom detects the user's human body activity at this time, and the sub-speaker sends the instruction for confirming the presence of the user to the master speaker. After receiving this instruction, the master speaker sends the volume increase control instruction (i.e. the second control play instruction) to the sub-speaker in the bedroom to control the playing sound of the sub-speaker to change from small to large.
  • In the present disclosure, the audio data sent by the master speaker based on the spatial identifier is acquired. The audio is played based on different spatial identifiers. Through the present disclosure, the master speaker can also control the change of volume of the sub-speaker playing audio according to the detection result of the sub-speaker in the subspace, so that the sub-speaker can achieve a gradual effect in volume and bring a good listening experience to the user.
  • Fig. 8 is a flow chart of playing audio based on audio data according to an illustrative embodiment. As shown in Fig. 8, playing audio based on the audio data includes the following steps.
  • In step S81, a system clock between the sub-speaker and the master speaker is calibrated.
  • In an embodiment of the present disclosure, Fig. 9 shows a schematic diagram of calibrating the system clock between the master speaker and the sub-speaker. As shown in Fig. 9, while the audio data is transmitted between the master speaker and the sub-speaker, the master speaker and the sub-speaker also need to have an interaction of clock information. The sub-speaker needs to calculate the system clock difference between itself and the master speaker, and adjust itself to the system clock consistent with the master speaker. According to its own clock, the sub-speaker sends time information to the master speaker through WiFi at time TB0, the master speaker receives the time information of the sub-speaker at time TA0 according to its own clock, then the master speaker sends time information to the sub-speaker through WiFi at time TA1, and the sub-speaker receives the time information of the master speaker at time TB1. The sub-speaker can calculate the system clock difference between itself and the master speaker by using the four time information of TB0, TA0, TA1 and TB1 as follows: T A 0 T B 0 = Δ + τ 0 and T B 1 T A 1 = Δ + τ 1 ,
    Figure imgb0003
    where Δ represents a system clock error between the sub-speaker and the master speaker, and τ 0 and τ 1 are WiFi transmission delays between the master speaker and the sub-speaker. Then, the system clock error can be calculated as: Δ = T A 0 T B 0 T B 1 T A 1 + τ 1 τ 0 2 .
    Figure imgb0004
    The sub-speaker compensates its own system clock with the clock difference Δ, and then it is calibrated to the system clock consistent with the master speaker. On the basis that all the speakers are aligned with the master speaker in clock, the audio frames at time T0 are played at time T0+1s, so that the synchronous play of all the speakers can be achieved.
  • In step S82, based on the calibrated system clock and the audio data, the audio is played synchronously with the master speaker.
  • In the present disclosure, the system clock between the sub-speaker and the master speaker is calibrated. Based on the calibrated system clock and the audio data, the audio is played synchronously with the master speaker. Through the present disclosure, based on the system clock criterion between the sub-speaker and the master speaker, the synchronous audio play of the respective speakers is realized, and a good listening experience is brought to the user.
  • Fig. 10 shows a schematic diagram of a speaker play control. As shown in Fig. 10, first, the space where each speaker is located is divided, and the division of the subspace can be realized by user configuration and automatic perception. The user configuration means that each time the user configures the network for the speaker, the user chooses the subspace in which the speaker is located. For example, the living room is chosen for speakers A and B, and the master bedroom is chosen for speakers C and D, so that speakers A and B mutually know that they are in a same subspace, and speakers C and D mutually know that they are in a same subspace. The automatic perception means that the user does not need to choose, and the speakers perceive each other through ultrasonic communication. For example, only speaker B can receive the ultrasonic information sent by speaker A, and speaker C and speaker D cannot receive it due to wall blocking, so that speaker B knows that it is in the same subspace with speaker A. Each speaker perceives in sequence, and then the subspaces where all the speakers are located can be finally divided. Speaker A is used as the master speaker, the information of the spaces where all the devices are located is finally summarized into speaker A, and the IP addresses or other identification IDs of the devices in the respective subspaces are stored in the list. In this case, the relationship between the speakers is established. In subspace 1, speaker M detects the user by ultrasonic technology. If speaker M detects that the user is in subspace 1, it sends an instruction to the master speaker to inform that there is the user in subspace 1. When the master speaker receives the instruction from speaker M, it continues to send the audio data to speaker M and speaker N in subspace 1. If speaker M does not detect that the user is in subspace 1, it sends an instruction to the master speaker to inform that there is no user in subspace 1. When the master speaker receives the instruction from speaker M, it stops sending the audio data to speaker M and speaker N in subspace 1, and closes the channels between the master speaker and speakers M and N. Through the present disclosure, the play function of the speaker in the room without the user is stopped. Thus, on the one hand, the unnecessary waste of electric energy is saved, and on the other hand, it saves the occupation of network resources, reduces the network load and helps to improve the stability of the system.
  • The embodiments of the present disclosure provide the speaker play control method, which is applied to the sub-speakers in the respective divided subspaces, and the sub-speakers communicate with the master speaker. When the sub-speaker detects that there is the user in the space through ultrasonic technology, it sends the instruction that there is the user to the master speaker. When the sub-speaker detects that there is no user in the space through ultrasonic technology, it sends the instruction that there is no user to the master speaker. The sub-speaker receives the audio data and the play control instruction from the master speaker, decodes the data and plays the decoded data, under the condition of informing the master speaker that there is the user in the space.
  • Based on the same concept, the embodiments of the present disclosure also provide a speaker play control device.
  • It can be understood that, in order to realize the above functions, the speaker play control device provided by the embodiments of the present disclosure includes corresponding hardware structures and/or software modules for executing various functions. In combination with the units and algorithm steps of various examples disclosed in the embodiments of the present disclosure, the embodiments of the present disclosure can be realized in the form of hardware or a combination of hardware and computer software. Whether a function is executed by hardware or in a manner of computer software driving hardware depends on the specific application and design constraints of the technical solution. Those skilled in the art can use different methods to realize the described functions for each specific application, but this realization should not be considered beyond the scope of the technical solution of the embodiment of the present disclosure.
  • Fig. 11 is a block diagram of a speaker play control device according to an illustrative embodiment. Referring to Fig. 11, the device 100 can be provided as the master speaker according to the above embodiments, and include a determining unit 101 and a playing unit 102.
  • The determining unit 101 is configured to acquire audio data from a server in response to enabling an audio play function, and determine a target subspace in subspaces to which a plurality of sub-speakers belong. The target subspace is a subspace where a user is currently located. The playing unit 102 is configured to send the audio data to the sub-speaker in the target subspace and control the sub-speaker in the target subspace to play audio based on the audio data.
  • In an embodiment, the determining unit 101 determines the target subspace in the subspaces to which the plurality of sub-speakers belong in the following ways: determining a user detection result, where the user detection result is determined by a set sub-speakers in the subspace based on detection of a human body activity of the user, and the user detection result includes a presence or an absence of the user; and determining the subspace to which the target sub-speaker whose detection result is the presence of the user belongs as the target subspace.
  • In an embodiment, the plurality of sub-speakers correspond to spatial identifiers, and the spatial identifier is used for identifying the subspace to which the sub-speaker belongs. The playing unit 102 sends the audio data to the sub-speaker in the target subspace and control the sub-speaker in the target subspace to play audio based on the audio data in the following ways: determining a target spatial identifier corresponding to the target subspace; sending the audio data to the sub-speaker in the target subspace identified by the target spatial identifier, and controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data.
  • In an embodiment, the target spatial identifier includes a first target spatial identifier and a second target spatial identifier, a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves. The playing unit 102 controls the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data in the following manners: sending a first control play instruction to the sub-speaker in the first target subspace, and controlling the sub-speaker in the first target subspace based on the first control play instruction to play audio based on the first control play instruction and the audio data, where the first control play instruction is used to control a playing sound of the sub-speaker to change from large to small; sending a second control play instruction to the sub-speaker in the second target subspace, and controlling the sub-speaker in the second target subspace based on the second control play instruction to play audio based on the second control play instruction and the audio data, where the second control play instruction is used to control the playing sound of the sub-speaker to change from small to large.
  • In an embodiment, the playing unit 102 is also used to: determine a non-target subspace, which is a subspace without the user to which each set sub-speaker belongs; and stop sending the audio data to the non-target subspace.
  • Fig. 12 is a block diagram of a speaker play control device according to an illustrative embodiment. Referring to Fig. 12, the device 200 can be provided as the sub-speaker according to the above embodiments, and include an acquiring unit 201 and a playing unit 202.
  • The acquiring unit 201 is configured to acquire audio data sent by a master speaker. The playing unit 202 is used to play audio based on the audio data.
  • In an embodiment, the sub-speaker is a set sub-speaker, and the set sub-speaker is used to detect whether there is a user in the subspace to which a target sub-speaker belongs. The playing unit 202 is also used to: detect a human body activity and determine a user detection result based on a human body activity detection result, where the user detection result includes a presence or an absence of a user; and send the user detection result to the master speaker.
  • In an embodiment, the sub-speaker corresponds to a spatial identifier, and the spatial identifier is used to identify the subspace to which the sub-speaker belongs. The acquiring unit 201 acquires the audio data sent by the master speaker in a following way: acquiring the audio data sent by the master speaker based on the spatial identifier.
  • In an embodiment, the playing unit 202 plays audio based on the audio data in the following ways: in response to that the spatial identifier is a first target spatial identifier, acquiring a first control play instruction sent by the master speaker, where a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and the first control play instruction is used to control a playing sound of the sub-speaker to change from large to small; based on the first control play instruction, playing audio based on the audio data in a manner that the playing sound changes from large to small.
  • In an embodiment, the playing unit 202 plays audio based on the audio data in the following ways: in response to that the spatial identifier is a second target spatial identifier, acquiring a second control play instruction sent by the master speaker, where the second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves, and the second control instruction is used to control the playing sound of the sub-speaker to change from small to large; and based on the second control play instruction, playing audio based on the audio data in a manner that the playing sound changes from small to large.
  • In an embodiment, the playing unit 202 plays audio based on the audio data in the following ways: calibrating a system clock between the sub-speaker and the master speaker; and based on the calibrated system clock and the audio data, playing audio synchronously with the master speaker.
  • With regard to the devices in the above embodiments, the specific way in which each module performs operations has been described in detail in the embodiments of the methods, and will not be described in detail here.
  • Fig. 13 is a block diagram of a device for speaker play control according to an illustrative embodiment. For example, a device 300 may be a mobile phone, a computer, a digital broadcasting terminal, a messaging device, a game console, a tablet device, a medical device, a fitness device, a personal digital assistant and the like.
  • Referring to Fig. 13, the device 300 may include one or more of the following components: a processing component 302, a memory 304, a power component 306, a multimedia component 308, an audio component 310, an input/output (I/O) interface 312, a sensor component 314, and a communication component 316.
  • The processing component 302 generally controls the overall operation of the device 300, such as operations associated with display, telephone call, data communication, camera operation and recording operation. The processing component 302 may include one or more processors 320 to execute instructions to complete all or part of the steps of the method described above. In addition, the processing component 302 can include one or more modules to facilitate the interaction between the processing component 302 and other components. For example, the processing component 302 can include a multimedia module to facilitate interaction between the multimedia component 308 and the processing component 302.
  • The memory 304 is configured to store various types of data to support operations in the device 300. Examples of the data include instructions for any application or method operating on the device 300, contact data, phone book data, messages, pictures, videos, and the like. The memory 304 can be realized by any type of volatile or nonvolatile memory device or their combination, such as static random access memory (SRAM), electrically erasable programmable read-only memory (EEPROM), erasable programmable read-only memory (EPROM), programmable read-only memory (PROM), read-only memory (ROM), magnetic memory, flash memory, magnetic disk or optical disk.
  • The power component 306 provides power for various components of the device 300. The power component 306 may include a power management system, one or more power sources, and other components associated with generating, managing and distributing power for the device 300.
  • The multimedia component 308 includes a screen that provides an output interface between the device 300 and the user. In some embodiments, the screen may include a liquid crystal display (LCD) and a touch panel (TP). If the screen includes a touch panel, the screen may be implemented as a touch screen to receive an input signal from the user. The touch panel includes one or more touch sensors to sense touch, sliding and gestures on the touch panel. The touch sensor may not only sense the boundary of a touch or sliding action, but also detect the duration and pressure related to the touch or sliding operation. In some embodiments, the multimedia component 308 includes a front camera and/or a rear camera. When the device 300 is in an operation mode, such as a shooting mode or a video mode, the front camera and/or the rear camera can receive external multimedia data. Each front camera and rear camera can be a fixed optical lens system or have a focal length and an optical zoom capability.
  • The audio component 310 is configured to output and/or input audio signals. For example, the audio component 310 includes a microphone (MIC) configured to receive external audio signals when the device 300 is in an operation mode, such as a call mode, a recording mode and a voice recognition mode. The received audio signal may be further stored in the memory 304 or transmitted via the communication component 316. In some embodiments, the audio component 310 further includes a speaker for outputting audio signals.
  • The I/O interface 312 provides an interface between the processing component 302 and peripheral interface modules, which can be keyboards, click wheels, buttons, etc. These buttons may include, but are not limited to, a home button, a volume button, a start button and a lock button.
  • The sensor assembly 314 includes one or more sensors for providing various aspects of the state evaluation for the device 300. For example, the sensor component 314 can detect the on/off state of the device 300, the relative positioning of components, such as the display and keypad of the device 300, the position change of the device 300 or a component of the device 300, the presence or absence of user contact with the device 300, the orientation or acceleration/deceleration of the device 300 and the temperature change of the device 300. The sensor assembly 314 may include a proximity sensor configured to detect the presence of a nearby object without any physical contact. The sensor assembly 314 may also include an optical sensor, such as a CMOS or CCD image sensor, for use in imaging applications. In some embodiments, the sensor assembly 314 may further include an acceleration sensor, a gyro sensor, a magnetic sensor, a pressure sensor or a temperature sensor.
  • The communication component 316 is configured to facilitate wired or wireless communication between the device 300 and other devices. The device 300 can access a wireless network based on communication standards, such as WiFi, 2G or 3G, or a combination thereof. In an illustrative embodiment, the communication component 316 receives a broadcast signal or broadcast related information from an external broadcast management system via a broadcast channel. In an illustrative embodiment, the communication component 316 further includes a near field communication (NFC) module to facilitate short-range communication. For example, the NFC module can be implemented based on radio frequency identification (RFID) technology, infrared data association (IrDA) technology, ultra-wideband (UWB) technology, Bluetooth (BT) technology and other technologies.
  • In an illustrative embodiment, the device 300 may be implemented by one or more application specific integrated circuits (ASIC), digital signal processors (DSP), digital signal processing devices (DSPD), programmable logic devices (PLD), field programmable gate arrays (FPGA), controllers, microcontrollers, microprocessors or other electronic components, for performing the above methods.
  • In an illustrative embodiment, there is also provided a non-transitory computer-readable storage medium including instructions, such as the memory 304 including instructions, and the instructions can be executed by the processor 320 of the device 300 to complete the above method. For example, the non-transitory computer-readable storage medium can be ROM, random access memory (RAM), CD-ROM, magnetic tape, floppy disk, optical data storage device, etc.
  • It can be understood that "a plurality of" in the present disclosure refers to two or more, and other quantifiers are similar. The term "and/or", which describes the relationship of related objects, means that there may be three kinds of relationships. For example, A and/or B can mean that A exists alone, A and B exist together, and B exists alone. The character "/" generally indicates that the former and latter objects have an OR relationship. The singular forms "a", "said" and "the" are also intended to include the plural forms, unless the context clearly indicates otherwise.
  • It is further understood that the terms "first" and "second" are used to describe various information, but the information should not be limited to these terms. These terms are only used to distinguish the same type of information from each other and do not indicate a specific order or importance. In fact, the expressions "first" and "second" can be used interchangeably. For example, without departing from the scope of the present disclosure, the first information may also be called the second information, and similarly, the second information may also be called the first information.
  • It can be further understood that unless otherwise specified, "connection" includes direct connection between two without other components, and indirect connection between two with other components.
  • It can be further understood that although the operations are described in a specific order in the drawings in the embodiments of the present disclosure, it should not be understood as requiring that these operations be performed in the specific order or serial order shown, or that all the operations shown should be performed to obtain the desired results. In certain circumstances, multitasking and parallel processing may be beneficial.
  • Other embodiments of the present disclosure will easily occur to those skilled in the art after considering the specification and practicing the invention disclosed herein. This application is intended to cover any variations, uses or adaptations of the present disclosure, which follow the general principles of the present disclosure and include the common sense or common technical means in the related art that are not disclosed in the present disclosure. The specification and embodiments are to be regarded as illustrative only, with the true scope and spirit of the present disclosure being indicated by the following claims.
  • It should be understood that the present disclosure is not limited to the precise structure described above and shown in the drawings, and various modifications and changes can be made without departing from the scope of the present disclosure. The scope of the present disclosure is limited only by the scope of the appended claims.

Claims (24)

  1. A speaker play control method, applied to a master speaker, the master speaker being configured to control a plurality of sub-speakers, the plurality of sub-speakers belonging to different subspaces of a same space, and the method comprising:
    in response to enabling an audio play function, acquiring audio data from a server, and determining a target subspace in the subspaces to which the plurality of sub-speakers belong, wherein the target subspace is a subspace where a user is currently located; and
    sending the audio data to a sub-speaker in the target subspace, and controlling the sub-speaker in the target subspace to play audio based on the audio data.
  2. The method according to claim 1, wherein the determining the target subspace in the subspaces to which the plurality of sub-speakers belong comprises:
    determining a user detection result, wherein the user detection result is determined by a set sub-speaker in the subspace based on detection of a human body activity of the user, and the user detection result comprises a presence or an absence of the user; and
    determining a subspace to which a target sub-speaker whose detection result is the presence of the user belongs as the target subspace.
  3. The method according to claim 1 or 2, wherein the plurality of sub-speakers correspond to spatial identifiers, and the spatial identifier is configured to identify the subspace to which the sub-speaker belongs;
    the sending the audio data to the sub-speaker in the target subspace and controlling the sub-speaker in the target subspace to play audio based on the audio data comprises:
    determining a target spatial identifier corresponding to the target subspace; and
    sending the audio data to the sub-speaker in the target subspace identified by the target spatial identifier, and controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data.
  4. The method according to claim 3, wherein the target spatial identifier comprises a first target spatial identifier and a second target spatial identifier, a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves;
    the controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data comprises:
    sending a first control play instruction to the sub-speaker in the first target subspace, and controlling the sub-speaker in the first target subspace based on the first control play instruction to play audio based on the first control play instruction and the audio data, wherein the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and
    sending a second control play instruction to the sub-speaker in the second target subspace, and controlling the sub-speaker in the second target subspace based on the second control play instruction to play audio based on the second control play instruction and the audio data, wherein the second control play instruction is configured to control the playing sound of the sub-speaker to change from small to large.
  5. The method according to claim 1, further comprising:
    determining a non-target subspace, wherein the non-target subspace is a subspace without the user to which each set sub-speaker belongs; and
    stopping sending the audio data to the non-target subspace.
  6. A speaker play control method, applied to a sub-speaker, a subspace to which the sub-speaker belongs being a target subspace, the target subspace being a subspace where a user is currently located, and the method comprising:
    acquiring audio data sent by a master speaker; and
    playing audio based on the audio data.
  7. The method according to claim 6, wherein the sub-speaker is a set sub-speaker, and the set sub-speaker is configured to detect whether there is the user in a subspace to which a target sub-speaker belongs, and the method further comprises:
    detecting a human body activity, and determining a user detection result based on a human body activity detection result, wherein the user detection result comprises a presence or an absence of the user; and
    sending the user detection result to the master speaker.
  8. The method according to claim 7, wherein the sub-speaker corresponds to a spatial identifier, and the spatial identifier is configured to identify the subspace to which the sub-speaker belongs; and
    the acquiring the audio data sent by the master speaker comprises:
    acquiring the audio data sent by the master speaker based on the spatial identifier.
  9. The method according to claim 8, wherein the playing audio based on the audio data comprises:
    in response to that the spatial identifier is a first target spatial identifier, acquiring a first control play instruction sent by the master speaker, wherein a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and
    based on the first control play instruction, playing audio based on the audio data in a manner that the playing sound changes from large to small.
  10. The method according to claim 8, wherein the playing audio based on the audio data comprises:
    in response to that the spatial identifier is a second target spatial identifier, acquiring a second control play instruction sent by the master speaker, wherein a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves, and the second control play instruction is configured to control a playing sound of the sub-speaker to change from small to large; and
    based on the second control play instruction, playing audio based on the audio data in a manner that the playing sound changes from small to large.
  11. The method according to claim 6, wherein the playing audio based on the audio data comprises:
    calibrating a system clock between the sub-speaker and the master speaker; and
    based on the calibrated system clock and the audio data, playing audio synchronously with the master speaker.
  12. A speaker play control device, comprising:
    a determining unit configured to acquire audio data from a server in response to enabling an audio play function, and determining a target subspace in subspaces to which a plurality of sub-speakers belong, wherein the target subspace is a subspace where a user is currently located; and
    a playing unit configured to send the audio data to a sub-speaker in the target subspace and control the sub-speaker in the target subspace to play audio based on the audio data.
  13. The device according to claim 12, wherein the determining unit determines the target subspace in the subspaces to which the plurality of sub-speakers belong in a following manner:
    determining a user detection result, wherein the user detection result is determined by a set sub-speaker in the subspace based on detection of a human body activity of the user, and the user detection result comprises a presence or an absence of the user; and
    determining a subspace to which a target sub-speaker whose detection result is the presence of the user belongs as the target subspace.
  14. The device according to claim 12 or 13, wherein the plurality of sub-speakers correspond to spatial identifiers, and the spatial identifier is configured to identify the subspace to which the sub-speaker belongs;
    the playing unit sends the audio data to the sub-speaker in the target subspace and controls the sub-speaker in the target subspace to play audio based on the audio data in a following manner:
    determining a target spatial identifier corresponding to the target subspace; and
    sending the audio data to the sub-speaker in the target subspace identified by the target spatial identifier, and controlling the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data.
  15. The device according to claim 14, wherein the target spatial identifier comprises a first target spatial identifier and a second target spatial identifier, wherein a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves;
    the playing unit controls the sub-speaker in the target subspace based on the target spatial identifier to play audio based on the audio data in a following manner;
    sending a first control play instruction to the sub-speaker in the first target subspace, and controlling the sub-speaker in the first target subspace based on the first control play instruction to play audio based on the first control play instruction and the audio data, wherein the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and
    sending a second control play instruction to the sub-speaker in the second target subspace, and controlling the sub-speaker in the second target subspace based on the second control play instruction to play audio based on the second control play instruction and the audio data, wherein the second control play instruction is configured to control the playing sound of the sub-speaker to change from small to large.
  16. The device according to claim 12, wherein the playing unit is further configured to:
    determine a non-target subspace, wherein the non-target subspace is a subspace without the user to which each set sub-speaker belongs; and
    stop sending the audio data to the non-target subspace.
  17. A speaker play control device, comprising:
    an acquiring unit configured to acquire audio data sent by a master speaker; and
    a playing unit configured to play audio based on the audio data.
  18. The device according to claim 17, wherein the sub-speaker is a set sub-speaker, and the set sub-speaker is configured to detect whether there is the user in a subspace to which a target sub-speaker belongs, and the playing unit is further configured to:
    detect a human body activity, and determining a user detection result based on a human body activity detection result, wherein the user detection result comprises a presence or an absence of the user; and
    send the user detection result to the master speaker.
  19. The device according to claim 18, wherein the sub-speaker corresponds to a spatial identifier, and the spatial identifier is configured to identify the subspace to which the sub-speaker belongs;
    the acquiring unit acquires the audio data sent by the master speaker in a following way:
    acquiring the audio data sent by the master speaker based on the spatial identifier.
  20. The device according to claim 19, wherein the playing unit plays audio based on the audio data in a following manner:
    in response to that the spatial identifier is a first target spatial identifier, acquiring a first control play instruction sent by the master speaker, wherein a first target subspace identified by the first target spatial identifier is a subspace where the user is before the user moves, and the first control play instruction is configured to control a playing sound of the sub-speaker to change from large to small; and
    based on the first control play instruction, playing audio based on the audio data in a manner that the playing sound changes from large to small.
  21. The device according to claim 19, wherein the playing unit plays audio based on the audio data in a following manner:
    in response to that the spatial identifier is a second target spatial identifier, acquiring a second control play instruction sent by the master speaker, wherein a second target subspace identified by the second target spatial identifier is a subspace where the user is after the user moves, and the second control play instruction is configured to control a playing sound of the sub-speaker to change from small to large; and
    based on the second control play instruction, playing audio based on the audio data in a manner that the playing sound changes from small to large.
  22. The device according to claim 17, wherein the playing unit plays audio based on the audio data in a following manner:
    calibrating a system clock between the sub-speaker and the master speaker; and
    based on the calibrated system clock and the audio data, playing audio synchronously with the master speaker.
  23. A speaker play control device, comprising:
    a processor; and
    a memory for storing instructions executable by the processor,
    wherein the processor is configured to perform a method according to any one of claims 1 to 5 or a method according to any one of claims 6 to 11.
  24. A computer-readable storage medium, wherein instructions are stored in the storage medium, and when the instructions in the storage medium are executed by a processor, the processor is enabled to perform a method according to any one of claims 1 to 5, or a method according to any one of claims 6 to 11.
EP22946317.9A 2022-06-17 2022-06-17 GAME CONTROL METHOD FOR A SOUND BOX, GAME CONTROL DEVICE FOR A SOUND BOX AND STORAGE MEDIUM Pending EP4542542A4 (en)

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
PCT/CN2022/099595 WO2023240636A1 (en) 2022-06-17 2022-06-17 Sound box playing control method, sound box playing control apparatus, and storage medium

Publications (2)

Publication Number Publication Date
EP4542542A1 true EP4542542A1 (en) 2025-04-23
EP4542542A4 EP4542542A4 (en) 2026-03-25

Family

ID=89192965

Family Applications (1)

Application Number Title Priority Date Filing Date
EP22946317.9A Pending EP4542542A4 (en) 2022-06-17 2022-06-17 GAME CONTROL METHOD FOR A SOUND BOX, GAME CONTROL DEVICE FOR A SOUND BOX AND STORAGE MEDIUM

Country Status (3)

Country Link
EP (1) EP4542542A4 (en)
CN (1) CN117597730A (en)
WO (1) WO2023240636A1 (en)

Family Cites Families (10)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
CN105828255A (en) * 2016-05-12 2016-08-03 深圳市金立通信设备有限公司 Method for optimizing pops and clicks of audio device and terminal
WO2020030303A1 (en) * 2018-08-09 2020-02-13 Fraunhofer-Gesellschaft zur Förderung der angewandten Forschung e.V. An audio processor and a method for providing loudspeaker signals
NO20181210A1 (en) * 2018-08-31 2020-03-02 Elliptic Laboratories As Voice assistant
US10506361B1 (en) * 2018-11-29 2019-12-10 Qualcomm Incorporated Immersive sound effects based on tracked position
CN111757218A (en) * 2019-03-29 2020-10-09 深圳市冠旭电子股份有限公司 Audio file playback method, distributed audio system and main speaker
CN113506568B (en) * 2020-04-28 2024-04-16 海信集团有限公司 Central control and intelligent equipment control method
CN112035086B (en) * 2020-08-19 2024-03-22 海尔优家智能科技(北京)有限公司 Audio playing method and device
CN114446290A (en) * 2020-11-02 2022-05-06 青岛海尔智能技术研发有限公司 Device response method, device, computer device and storage medium
CN112449285A (en) * 2020-12-15 2021-03-05 广州市登宝路电器有限公司 Method and system for synchronously playing multiple sound boxes
CN114299951A (en) * 2021-12-31 2022-04-08 联想(北京)有限公司 A control method and device

Also Published As

Publication number Publication date
CN117597730A (en) 2024-02-23
EP4542542A4 (en) 2026-03-25
WO2023240636A1 (en) 2023-12-21

Similar Documents

Publication Publication Date Title
US11425769B2 (en) Wireless communication method, terminal, audio component, device, and storage medium
US11469962B2 (en) Method and apparatus for configuring information of indicating time-frequency position of SSB, and method and apparatus for determining time-frequency position of SSB
US10306437B2 (en) Smart device grouping system, method and apparatus
JP6314286B2 (en) Audio signal optimization method and apparatus, program, and recording medium
CN108399917B (en) Speech processing method, apparatus and computer readable storage medium
WO2017071070A1 (en) Speech control method and apparatus for smart device, control device and smart device
US11646856B2 (en) Timing configuration method and apparatus
US10827455B1 (en) Method and apparatus for sending a notification to a short-range wireless communication audio output device
CN105262452A (en) Method and apparatus for adjusting volume, and terminal
CN103023866A (en) Easy sharing of wireless audio signals
CN105532634A (en) Ultrasonic wave mosquito repel method, device and system
CN113015065A (en) Connection method, device, equipment and computer storage medium
US11197081B2 (en) Method for determining configuration parameter and earphone
CN105451056B (en) Audio and video synchronization method and device
EP3264130A1 (en) Method and apparatus for screen state switching control
CN109121047B (en) Stereo realization method of double-screen terminal, terminal and computer readable storage medium
KR20170023753A (en) Method and apparatus for collecting sound of monitoring screen
JP2022507520A (en) Broadcasting, receiving method and equipment of the configuration information of the synchronization signal block
CN104682908A (en) Method and device for controlling volume
CN104159283A (en) Method and device for controlling message transmission
CN108206884B (en) Terminal, adjusting method for communication signal transmitted by terminal and electronic equipment
CN114639383B (en) Device wake-up methods, apparatus, electronic devices and media
CN111123716A (en) Remote control method, remote control device, and computer-readable storage medium
EP4542542A1 (en) Sound box playing control method, sound box playing control apparatus, and storage medium
CN106451824A (en) Bluetooth earphone charging method, Bluetooth earphone charging device and Bluetooth earphone

Legal Events

Date Code Title Description
STAA Information on the status of an ep patent application or granted ep patent

Free format text: STATUS: THE INTERNATIONAL PUBLICATION HAS BEEN MADE

PUAI Public reference made under article 153(3) epc to a published international application that has entered the european phase

Free format text: ORIGINAL CODE: 0009012

STAA Information on the status of an ep patent application or granted ep patent

Free format text: STATUS: REQUEST FOR EXAMINATION WAS MADE

17P Request for examination filed

Effective date: 20240730

AK Designated contracting states

Kind code of ref document: A1

Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC MK MT NL NO PL PT RO RS SE SI SK SM TR

DAV Request for validation of the european patent (deleted)
DAX Request for extension of the european patent (deleted)
REG Reference to a national code

Ref country code: DE

Ref legal event code: R079

Free format text: PREVIOUS MAIN CLASS: G10L0015220000

Ipc: H04S0007000000

A4 Supplementary search report drawn up and despatched

Effective date: 20260223

RIC1 Information provided on ipc code assigned before grant

Ipc: H04S 7/00 20060101AFI20260217BHEP

Ipc: H04R 3/12 20060101ALI20260217BHEP

Ipc: G10L 15/22 20060101ALI20260217BHEP