WO2014190824A1 - Method, system and computer storage medium for detecting an audio input interface - Google Patents

Method, system and computer storage medium for detecting an audio input interface Download PDF

Info

Publication number
WO2014190824A1
WO2014190824A1 PCT/CN2014/075663 CN2014075663W WO2014190824A1 WO 2014190824 A1 WO2014190824 A1 WO 2014190824A1 CN 2014075663 W CN2014075663 W CN 2014075663W WO 2014190824 A1 WO2014190824 A1 WO 2014190824A1
Authority
WO
WIPO (PCT)
Prior art keywords
audio input
interface
input signals
interfaces
acquiring
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/CN2014/075663
Other languages
French (fr)
Inventor
Hong Liu
Chao Peng
Yuanjiang Peng
Xingping LONG
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Tencent Technology Shenzhen Co Ltd
Original Assignee
Tencent Technology Shenzhen Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Tencent Technology Shenzhen Co Ltd filed Critical Tencent Technology Shenzhen Co Ltd
Publication of WO2014190824A1 publication Critical patent/WO2014190824A1/en
Priority to US14/683,103 priority Critical patent/US9886238B2/en
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Classifications

    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F3/00Input arrangements for transferring data to be processed into a form capable of being handled by the computer; Output arrangements for transferring data from processing unit to output unit, e.g. interface arrangements
    • G06F3/16Sound input; Sound output
    • G06F3/167Audio in a user interface, e.g. using voice commands for navigating, audio feedback
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04RLOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
    • H04R3/00Circuits for transducers
    • H04R3/005Circuits for transducers for combining the signals of two or more microphones
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/78Detection of presence or absence of voice signals
    • G10L2025/783Detection of presence or absence of voice signals based on threshold decision
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04MTELEPHONIC COMMUNICATION
    • H04M3/00Automatic or semi-automatic exchanges
    • H04M3/42Systems providing special services or facilities to subscribers
    • H04M3/56Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities
    • H04M3/561Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities by multiplexing
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04MTELEPHONIC COMMUNICATION
    • H04M3/00Automatic or semi-automatic exchanges
    • H04M3/42Systems providing special services or facilities to subscribers
    • H04M3/56Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities
    • H04M3/568Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities audio processing specific to telephonic conferencing, e.g. spatial distribution, mixing of participants
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04MTELEPHONIC COMMUNICATION
    • H04M3/00Automatic or semi-automatic exchanges
    • H04M3/42Systems providing special services or facilities to subscribers
    • H04M3/56Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities
    • H04M3/568Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities audio processing specific to telephonic conferencing, e.g. spatial distribution, mixing of participants
    • H04M3/569Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities audio processing specific to telephonic conferencing, e.g. spatial distribution, mixing of participants using the instant speaker's algorithm
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04RLOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
    • H04R27/00Public address systems

Definitions

  • the present disclosure relates to the field of audio processing , and more particularly to a method and a device for detecting an audio input interface.
  • an object of the present disclosure is to provide a method for detecting an audio input interface, by means of detecting the input data of every one of audio input interfaces, the audio input interface connected to the microphone, into which a voice signal is being input, can be effectively identified, which does not need the user to switch manually and is very convenient.
  • the present disclosure is realized by the following technical scheme:
  • a method for detecting an audio input interface comprises:
  • the present disclosure is to provide a system for detecting an audio input interface, comprising:
  • an input detecting module configured to acquire input signals of every one of audio input interfaces
  • an energy detecting module configured to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
  • an interface identification acquiring module configured to add said interface identification acquired herein into an identification sequence in an order by acquiring time
  • an identifying module configured to identify the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
  • input signals of every one of audio input interfaces are acquired at a preset time interval, energy values of said input signals of every one of said audio input interfaces are detected, and an interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired.
  • Each of the interface identifications acquired is added to a preset identification sequence in an order by acquiring time. If the same interface identification is saved more often in said identification sequence, it means that, there are more such situations that the energy value of the input signal of the audio input interface corresponding to said i nterface identification is the maximum, and the audio input interface corresponding to said interface identification will be identified as a valid audio input interface.
  • the present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, and the user does not need to switch manually, which is very convenient, moreover, the influence of the noise on the result of identifying is reduced, thereby the result of identifying is more accurate .
  • FIG. 1 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the first embodiment of the present disclosure ;
  • FIG. 2 is a schematic flow diagram illustrating the method for detecting an audio input interface according tothe second embodiment of the present disclosure
  • FIG. 3 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the third embodiment of the present disclosure ;
  • FIG. 4 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the fourth embodiment of the present disclosure
  • FIG. 5 is a structure diagram illustrating the system for detecting an audio input interface according to the first embodiment of the present disclosure
  • FIG. 6 is a structure diagram illustrating the system for detecting an audio input interface according to the third embodiment of the present disclosure.
  • FIG. 7 is a structure diagram illustrating the system for detecting an audio input interface according to the fourth embodiment of the present disclosure.
  • Fig. 1 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the first embodiment of the present disclosure
  • the method for detecting an audio input interface of said embodiment comprises following steps:
  • S101 acquiring input signals of every one of audio input interfaces ;
  • S102 acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
  • S104 identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
  • input signals of every one of audio input interfaces are acquired at a preset time interval, energy values of said input signals of every one of said audio input interfaces are detected, and an interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired.
  • Each of the interface identifications acquired is added to a preset identification sequence in an order by acquiring time. If the same interface identification is saved more often in said identification sequence, it means that, there are more such situations that the energy value of the input signal of the audio input interface corresponding to said interface identification is the maximum, and the audio input interface corresponding to said interface identification will be identified as a valid audio input interface.
  • the method of the present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, and the user does not need to switch manually, which is very convenient, moreover, the influence of the noise on the result of identifying is reduced, thereby the result of identifying is more accurate.
  • the input signals of every one of audio input interfaces can be acquired through monitoring for and collecting data from each of saidaudio input interfaces.
  • the input signals of every one of said audio input interfaces comprise input signals from a microphone hardware-connected to any one of said audio input interfaces, noise signals, and so on.
  • all of the audio input interfaces can be enumerated by means of calling a function of Dsound API (Direct Sound Capture Enumerate( )).
  • the input data of every one of theaudio input interfaces is acquired by means of collecting the input signals of every one of said audio input interfaces.
  • the parameter of each of the audio input interfaces is preset to unify the audio collecting format for each of the audio input interfaces, such as using an audio collecting format with mono-channel, and 44.1 KHz of sampling rate.
  • input signals of every one of audio input interfaces are acquired simultaneously at a preset time interval, which comprises following sub-steps:
  • S1011 simultaneously collecting input signals of every one of said audio input interfaces, and encapsulating said input signals of every one of said audio input interfaces collected at the same time into a frame of detection data.
  • S1012 de-interleaving each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces. Moreover, each of frames of said detection data can be saved; one of said frames of detection data is extracted every preset frames and further de-interleaved so as to acquire the input signals of every one of said audio input interfaces contained in said frame of detection data .
  • each of the input signals is collected in an unit of 20 milliseconds, and then signals of every one of the input signals are put in a buffer with a M*20(M represents the number of the devices enumerated)milliseconds length, the corresponding data is acquired from said buffer and encapsulated.
  • the input signals (such as N channels of input signals, N is a nature number) of every one of audio input interfaces are encapsulated in respective frames of detection data.
  • the purpose of sampling said input signals at a preset time interval can be realized.
  • the input signals of every one of audio input interfaces can be re-acquired, which is very convenient.
  • each of the input signals is pre-processed so as to ensure the accuracy of the result of the detection.
  • the pre-processing comprises high-pass filtering, filtering certain frequency interferences, noise suppression, and so on, so as to reduce the influence of the noise on the detection of the input signals.
  • each energy value represents the intensity of a respective input signal.
  • the energy value of an input signal is the maximum among the energy values, means that the intensity of the input signal is the strongest, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user.
  • the interface identification of said audio input interface is acquired to identify the audio input interface of the input signal whose intensity is the strongest in the present detection.
  • said interface identification acquired herein is added into an identification sequence in an order by acquiring time.
  • Said identification sequence may be created in a buffer or in other types of memories, so that it is easy to be accessed.
  • said identification sequence follows the rule of first-in-and-first-out in an order by acquiring time, and the quantity of the interface identifications saved each time is less than or equal to a preset quantity. That is, if the quantity of interface identifications saved in said identification sequence reaches said preset quantity, whenever a new interface identification is added, an interface identification of which the acquiring time (in another word, the save time) is the earliest will be discarded.
  • the quantity of theinterfaceidentifications saved in saididentification sequence is maintained in said preset quantity, such as 25, and the most recently acquired 25 of interface identifications will always be saved.
  • the interface identifications saved in the identification sequence may be the same, that is, said identification sequence may contain multiple of the same interface identification.
  • S104 identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
  • the more often the same interface identifications is saved in said identification sequence means that the more often the audio frequency with the maximum value is input into the audio input interface corresponding to said interface identification, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user, and the correct chance will be high to identify said audio input interface corresponding to said interface identification as a valid audio input interface.
  • the accuracy of identifying will be further increased.
  • said audio input interface may be automatically matched to an audio software in the background for processing.
  • the input signal of said audio input interface may be subject to processing such as filtering and so on, before it is output into said audio software in the background for processing.
  • a user interface may be further displayed, on which the audio input interfaceidentified currently or the microphone or other audio input devices connected to said audio input interface will be shown.
  • the start and the end of the method for detecting an audio input interface of the present disclosure may be triggered through many different modes, for example, it may be set to start the detection when it is detected that the microphone is inserted, or it may be set to start the detection when a start instruction is received; also for example, it may be set to end the detection when the right input interface is detected, or it may be set to end the detection when the microphone is removed, or it may be set to end the detection when an end instruction is received, and so on.
  • Fig. 2 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the second embodiment of the present disclosure
  • the step S102 is specified as following sub-steps:
  • the method for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 1 in that, prior to said acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values (S1026), the method further comprises step S1024.
  • the interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired only when it is judged that said maximum energy value among the energy values of every one of said input signals is greater than or equal to a preset energy value, otherwise, the input signals of every one of said audio input interfaces are judged as all invalid.
  • the audio input with the maximum energy value is created by a noise, said audio input will be judged as an invalid signal as long as the intensity of noise is less than said preset energy value, and theresult of the identificationwill not be affected, the influence of the noise on the result of the identification is effectively reduced.
  • Fig. 3 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the third embodiment of the present disclosure
  • the method for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 1 in that, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step S105:
  • step S102 judging the input signals of every one of audio input interfaces as valid, and going to step S102 to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
  • Said VAD detection is so called voice activity detection, which can effectively detect the activity of the input signals, identify the input signal which may be the audio input, and increase the speed of identifying the activities audio input interface. If all results of the VAD detection are 0, that means every one of said audio input interfaces are currently in a muted state; if at least one of results of said audio input interfaces is 1, then at least one of said audio input interfaces has an audio input, and the input signals of every one of audio input interfaces collected herein can be judged as valid, and then the next step (acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values) is executed.
  • Fig. 4 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the fourth embodiment of the present disclosure
  • the method for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 1 in that, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step S106:
  • step S102 judging every one of said input signals as valid, and going to step S102 to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
  • the so called signal-noise ratio is a ratio of a normal sound signal to a noise signal (power) with no-sound signal, and it is often represented in the unit of dB.
  • the signal-noise ratios of every one of audio input signals are detected, and the input signals of the audio input interface are judged as valid only when at least one of the signal-noise ratios of said audio input interfaces is no less than the preset value of signal-noise ratio, otherwise, the input signals of the audio input interface are judged as invalid.
  • the influence of the noise on identifying an active audio input interface is reduced, and the accuracy of identifying is improved.
  • any two or any three of steps S1021, S105 and S106 can be selected and executed in combination simultaneously, thereby further improving the accuracy and efficiency of identifying.
  • Said post-processing comprises: executing a self-adaptive microphone volume adjustment to said valid audio input interface, that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state; executing a signal-noise ratio detection, starting a noise suppression according to the result of the detection, and so on.
  • a self-adaptive microphone volume adjustment to said valid audio input interface that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state
  • AGC Automatic Gain Control
  • the user does not need to participate in setting the device, but the device will automatically select the microphone, to which the user is inputting voice signals, and the problem of no sound will never happen, and the corresponding result of the configuration will be displayed in a user interface.
  • the user simply needs to speak to the microphone he wants, then the automatically switching of the audio input interfaces will be achieved, without the need of manual setting. If the microphone is broken and no sound is collected, the device can also automatically switch off the corresponding audio input interface, without the need of manual setting.
  • the user who doesn't know how to set the microphone is relieved from the difficulty, which is a convenience for the user.
  • said system for detecting an audio input interface of present disclosure comprises:
  • an input detecting module 11 configured to acquire input signals of every one of audio input interfaces
  • an energy detecting module 12 configured to acquire signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
  • an interface identification acquiring module 13 configured to add said interface identification acquired herein into an identification sequence in an order by acquiring time;
  • an identifying module 14 configured to the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
  • input signals of every one of audio input interfaces are acquired at a preset time interval, energy values of said input signals of every one of said audio input interfaces are detected, and an interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired.
  • Each of the interface identifications acquired is added to a preset identification sequence in an order by acquiring time. If the same interface identification is saved more often in said identification sequence, it means that, there are more such situations that the energy value of the input signal of the audio input interface corresponding to said interface identification is the maximum, and the audio input interface corresponding to said interface identification will be identified as a valid audio input interface.
  • the system of the present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, and the user does not need to switch manually, which is very convenient, moreover, the influence of the noise on the result of identifying is reduced, thereby the result of identifying is more accurate.
  • the input signals of every one of audio input interfaces can be acquired through monitoring for and collecting data from each of said audio input interfaces.
  • the input signals of every one of said audio input interfaces comprises input signals from a microphone hardware-connected to any one of said audio input interfaces, noise signals, and so on.
  • all of the audio input interfaces can be enumerated by means of calling a function of Dsound API (Direct Sound Capture Enumerate( )).
  • the input data of every one of the audio input interfaces is acquired by means of collecting the input signals of every one of said audio input interfaces.
  • the parameter of each of the audio input interfaces is preset to unify the audio collecting format for each of the audio input interfaces, such as using an audio collecting format with mono-channel, and 44.1 KHz of sampling rate.
  • said input detecting module 11 comprises following sub-modules:
  • a collecting unit configured to simultaneously collecting input signals of every one of said audio input interfaces ;
  • an encapsulating unit configured to encapsulate said input signals of every one of said audio input interfaces being collected at the same time into a frame of detection data
  • an extracting unit configured to de-interleave each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces contained in said frame of detection data.
  • said input detecting module 11 further comprises a saving unit, configured to save each of frames of said detection data; said extracting unit is further configured to extract one of said frames of detection data every preset frames, and then de-interleave said frame extracted.
  • each of the input signals is collected in an unit of 20 milliseconds, and then signals of every one of the input signals are put in a buffer with a M*20 ( M represents the number of the devices enumerated ) milliseconds length, the corresponding data is acquired from said buffer and encapsulated.
  • the input signals (such as N channels of input signals, N is a nature number) of every one of audio input interfaces are encapsulated in respective frames of detection data.
  • the purpose of sampling said input signals at a preset time interval can be realized.
  • the input signals of every one of audio input interfaces can be re-acquired, which is very convenient.
  • each of the input signals is pre-processed so as to ensure the accuracy of the result of the detection.
  • the pre-processing comprises high-pass filtering, filtering certain frequency interferences, noise suppression, and so on, so as to reduce the influence of the noise on the detection of the input signals.
  • each energy value represents the intensity of a respective input signal.
  • the energy value of an input signal is the maximum among the energy values, means that the intensity of the input signal is the strongest, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user.
  • the interface identification of said audio input interface is acquired to identify the audio input interface of the input signal whose intensity is the strongest in the present detection.
  • said interface identification acquired herein is added into an identification sequence in an order by acquiring time.
  • Said identification sequence may be created in a buffer or in other types of memories, so that it is easy to be accessed.
  • said identification sequence follows the rule of first-in-and-first-out in an order by acquiring time, and the quantity of the interface identifications saved each time is less than or equal to a preset quantity. That is, if the quantity of interface identifications saved in said identification sequence reaches said preset quantity, whenever a new interface identification is added, an interface identification of which the acquiring time (in another word, the save time) is the earliest will be discarded.
  • the quantity of the interface identifications saved in said identification sequence is maintained in said preset quantity, such as 25, and the most recently acquired 25 of interface identifications will always be saved.
  • the interface identifications saved in the identification sequence may be the same, that is, said identification sequence may contain multiple of the same interface identification.
  • identifying module 14 it is configured to identify the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
  • the more often the same interface identifications is saved in said identification sequence means that the more often the audio frequency with the maximum value is input into the audio input interface corresponding to said interface identification, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user, and the correct chance will be high to identify said audio input interface corresponding to said interface identification as a valid audio input interface. And with the increasing of the quantity of the interface identifications saved in said identification sequence, the accuracy of identifying will be further increased.
  • said audio input interface may be automatically matched to an audio software in the background for processing.
  • the input signal of said audio input interface may be subject to processing such as filtering and so on, before it is output into said audio software in the background for processing.
  • a user interface may be further displayed, on which the audio input interface identified currently or the microphone or other audio input devices connected to said audio input interface will be shown.
  • the start and the end of the system for detecting an audio input interface of the present disclosure may be triggered through many different modes, for example, it may be set to start the detection when it is detected that the microphone is inserted, or it may be set to start the detection when a start instruction is received; also for example, it may be set to end the detection when the right input interface is detected, or it may be set to end the detection when the microphone is removed, or it may be set to end the detection when an end instruction is received, and so on.
  • the present disclosure further provides a second embodiment of the system for detecting an audio input interface, which is mainly different from the first embodiment as shown in Fig. 5 in that, prior to acquiring an interface identification of said audio input interface of which the energy value is the maximum among the energy values, said interface identification acquiring module 13 further judges whether said maximum energy value among the energy values of every one of said input signals is no less than (greater than or equal to) a preset energy value; if it is, then acquires the interface identification of the audio input interface of which the energy value is the maximum among the energy values; if it is not, then judges the input signals of every one of the audio input interfaces as all invalid, and re-acquires input signals of every one of said input interfaces;
  • said interface identification acquiring module 13 acquires the interface identification of the audio input interface of which the energy value is the maximum among the energy values only when it is judged that said maximum energy value among the energy values of every one of said input signals is greater than or equal to a preset energy value, otherwise, the input signals of every one of said audio input interfaces are judged as all invalid.
  • said audio input with the maximum energy value is created by a noise, said audio input will be judged as an invalid signal as long as the intensity of noise is less than said preset energy value, and the result of the identification will not be affected, the influence of the noise on the result of the identification is effectively reduced.
  • Fig. 6 is a structure diagram illustrating the system for detecting an audio input interface according to the third embodiment of the present disclosure
  • the system for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 5 in that, the system further comprises a VAD detecting module 15;
  • Said VAD detecting module 15 is configured to execute VAD detection on the input signals of every one of said audio input interfaces;
  • Said VAD detection is so called voice activity detection, which can effectively detect the activity of the input signals, identify the input signal which may be the audio input, and increase the speed of identifying the activities audio input interface. If all results of the VAD detection are 0, that means every one of said audio input interfaces are currently in a muted state; if at least one of results of said audio input interfaces is 1, then at least one of said audio input interfaces has an audio input, and the input signals of every one of audio input interfaces collected herein can be judged as valid, and then the next step (acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values) is executed.
  • the system is mainly different from the first embodiment as shown in Fig. 5 in that, the system further comprises a signal-noise ratio detecting module 16, which is configured to acquire signal-noise ratios of the input signals of every one of said audio input interfaces; if each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then judging the input signals of every one of said audio input interfaces as invalid, and re-acquiring input signals of every one of said audio input interfaces; if at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then judging every one of said input signals as valid, said energy detecting module 12 acquires energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the
  • the so called signal-noise ratio is a ratio of a normal sound signal to a noise signal (power) with no-sound signal, and it is often represented in the unit of dB.
  • the signal-noise ratios of every one of audio input signals are detected, and the input signals of the audio input interface are judged as valid only when at least one of the signal-noise ratios of said audio input interfaces is no less than the preset value of signal-noise ratio, otherwise, the input signals of the audio input interface are judged as invalid.
  • the influence of the noise on identifying an active audio input interface is reduced, and the accuracy of identifying is improved.
  • any two or any three of said interface identification acquiring module 13, said VAD detecting module 15 and said signal-noise ratio detecting module 16 can be selected and adopted in combination, thereby further improving the accuracy and efficiency of identifying.
  • post-processing is further executed to the audio input interface identified by adjusting parameters related to the device so as to make the microphone connected to said device in the best working state.
  • Said post-processing comprises: executing a self-adaptive microphone volume adjustment to said valid audio input interface, that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state; executing a signal-noise ratio detection, starting a noise suppression according to the result of the detection, and so on.
  • a self-adaptive microphone volume adjustment to said valid audio input interface that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state
  • AGC Automatic Gain Control
  • the user does not need to participate in setting the device, but the device will automatically select the microphone, to which the user is inputting voice signals, and the problem of no sound will never happen, and the corresponding result of the configuration will be displayed in a user interface.
  • the user simply needs to speak to the microphone he wants, then the automatically switching of the audio input interfaces will be achieved, without the need of manual setting. If the microphone is broken and no sound is collected, the device can also automatically switch off the corresponding audio input interface, without the need of the user's manual setting.
  • the user who doesn't know how to set the microphone is relieved from the difficulty, which is a convenience for the user.
  • the person skilled in the art can understand that, all of or part of the processes and the corresponding system implementing the embodiments mentioned above, may be achieved by means of relevant hardware commanded by computer programs, the computer programs may be saved in the computer readable storage medium, and they may comprise the processes of embodiments of the respective methods and systems mentioned above when the programs are executed.
  • the storage medium may be a disk or CD or read-only memory or random access memory, and so on.

Landscapes

  • Engineering & Computer Science (AREA)
  • General Health & Medical Sciences (AREA)
  • Health & Medical Sciences (AREA)
  • Physics & Mathematics (AREA)
  • Theoretical Computer Science (AREA)
  • Acoustics & Sound (AREA)
  • Signal Processing (AREA)
  • Otolaryngology (AREA)
  • Multimedia (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • General Engineering & Computer Science (AREA)
  • General Physics & Mathematics (AREA)
  • Circuit For Audible Band Transducer (AREA)
  • User Interface Of Digital Computer (AREA)
  • Telephone Function (AREA)

Abstract

Provided is a method for detecting an audio input interface, comprising: acquiring input signals of every one of audio input interfaces; acquiring energy values of said input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values; adding said interface identification acquired herein into an identification sequence in an order by acquiring time; and identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface. Also provided is a system for detecting an audio input interface. The present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, so that the user does not need to switch manually.

Description

Method , System and Computer Storage Medium for Detecting an Audio Input Interface
【FIELD OF THE INVENTION】
This application claims priority to Chinese Application No. 201310202043.3, entitled 'Method and System for Detecting an Audio Input Interface', filed on May 27, 2013, which is hereby incorporated by reference in its entirety.
The present disclosure relates to the field of audio processing , and more particularly to a method and a device for detecting an audio input interface.
【BACKGROUND OF THE INVENTION】
Along with the popularity ofvoice software, it is becoming more accepted by most computer users, and has been an indispensable part in people's daily life gradually. The existing computer device often provides an option for choosing an audio input interface, which needs the user to manually switch to choose among different audio input interfaces, however, the switching method herein requires the user to manually try to choose each audio input interface one by one until a voice signal is heard, which is very inconvenient. Moreover, the user often makes misconnection due to not knowing the correct audio input interface, as a result, the correct voice input can not be acquired.
【SUMMARY OF THE INVENTION】
In view of the defects existing in conventional method and device mentioned above that, the microphone, to which a user is inputting voice signals, can not be identified automatically, but the user needs to manually switch to choose among audio input interfaces one by one, which is very inconvenient, an object of the present disclosure is to provide a method for detecting an audio input interface, by means of detecting the input data of every one of audio input interfaces, the audio input interface connected to the microphone, into which a voice signal is being input, can be effectively identified, which does not need the user to switch manually and is very convenient.
In one aspect, the present disclosure is realized by the following technical scheme:
A method for detecting an audio input interface comprises:
acquiring input signals of every one of audio input interfaces;
acquiring energy values of said input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
adding said interface identification acquired herein into an identification sequence in an order by acquiring time; and
identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
In another aspect, the present disclosure is to provide a system for detecting an audio input interface, comprising:
an input detecting module, configured to acquire input signals of every one of audio input interfaces;
an energy detecting module, configured to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
an interface identification acquiring module, configured to add said interface identification acquired herein into an identification sequence in an order by acquiring time; and
an identifying module, configured to identify the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
According to the method and system for detecting an audio input interface of the present disclosure, input signals of every one of audio input interfaces are acquired at a preset time interval, energy values of said input signals of every one of said audio input interfaces are detected, and an interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired. Each of the interface identifications acquired is added to a preset identification sequence in an order by acquiring time. If the same interface identification is saved more often in said identification sequence, it means that, there are more such situations that the energy value of the input signal of the audio input interface corresponding to said i nterface identification is the maximum, and the audio input interface corresponding to said interface identification will be identified as a valid audio input interface. The present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, and the user does not need to switch manually, which is very convenient, moreover, the influence of the noise on the result of identifying is reduced, thereby the result of identifying is more accurate .
【BRIEF DESCRIPTION OF THE DRAWINGS】
FIG. 1 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the first embodiment of the present disclosure ;
FIG. 2 is a schematic flow diagram illustrating the method for detecting an audio input interface according tothe second embodiment of the present disclosure;
FIG. 3 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the third embodiment of the present disclosure ;
FIG. 4 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the fourth embodiment of the present disclosure;
FIG. 5 is a structure diagram illustrating the system for detecting an audio input interface according to the first embodiment of the present disclosure;
FIG. 6 is a structure diagram illustrating the system for detecting an audio input interface according to the third embodiment of the present disclosure;
FIG. 7 is a structure diagram illustrating the system for detecting an audio input interface according to the fourth embodiment of the present disclosure.
【DETAILED DESCRIPTION OF THE EMBODIMENTS】
In order to make the purpose, technical solutions and advantages of the present disclosure to be understood more clearly, the present disclosure will be described in further details with the accompanying drawings and the following embodiments. It should be understood that the specific embodiments described herein are merely examples to illustrate the disclosure, not to limit the present disclosure.
As shown in Fig. 1, which is a schematic flow diagram illustrating the method for detecting an audio input interface according to the first embodiment of the present disclosure, the method for detecting an audio input interface of said embodiment comprises following steps:
S101: acquiring input signals of every one of audio input interfaces ;
S102: acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
S103: adding said interface identification acquired herein into a preset identification sequence in an order by acquiring time ;
S104: identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
According to the method for detecting an audio input interface of the present disclosure, input signals of every one of audio input interfaces are acquired at a preset time interval, energy values of said input signals of every one of said audio input interfaces are detected, and an interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired. Each of the interface identifications acquired is added to a preset identification sequence in an order by acquiring time. If the same interface identification is saved more often in said identification sequence, it means that, there are more such situations that the energy value of the input signal of the audio input interface corresponding to said interface identification is the maximum, and the audio input interface corresponding to said interface identification will be identified as a valid audio input interface. The method of the present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, and the user does not need to switch manually, which is very convenient, moreover, the influence of the noise on the result of identifying is reduced, thereby the result of identifying is more accurate.
Wherein, with respect to step S101 , the input signals of every one of audio input interfaces can be acquired through monitoring for and collecting data from each of saidaudio input interfaces. The input signals of every one of said audio input interfaces comprise input signals from a microphone hardware-connected to any one of said audio input interfaces, noise signals, and so on. In one embodiment, all of the audio input interfaces can be enumerated by means of calling a function of Dsound API (Direct Sound Capture Enumerate( )).
After all of the audio input interfaces of the device having been detected, the input data of every one of theaudio input interfaces is acquired by means of collecting the input signals of every one of said audio input interfaces. Preferably, the parameter of each of the audio input interfaces is preset to unify the audio collecting format for each of the audio input interfaces, such as using an audio collecting format with mono-channel, and 44.1 KHz of sampling rate. By unifying the audio collecting format of each of the audio input interfaces, a large amount of computational load can be reduced in the post processing of the input signals, and the speed of identifying the microphone can be increased.
In one embodiment, input signals of every one of audio input interfaces are acquired simultaneously at a preset time interval, which comprises following sub-steps:
S1011: simultaneously collecting input signals of every one of said audio input interfaces, and encapsulating said input signals of every one of said audio input interfaces collected at the same time into a frame of detection data.
S1012: de-interleaving each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces. Moreover, each of frames of said detection data can be saved; one of said frames of detection data is extracted every preset frames and further de-interleaved so as to acquire the input signals of every one of said audio input interfaces contained in said frame of detection data .
In one embodiment, whenencapsulatingthe input signals of every one of said audio input interfaces, each of the input signals is collected in an unit of 20 milliseconds, and then signals of every one of the input signals are put in a buffer with a M*20(M represents the number of the devices enumerated)milliseconds length, the corresponding data is acquired from said buffer and encapsulated.
In this way, the input signals (such as N channels of input signals, N is a nature number) of every one of audio input interfaces are encapsulated in respective frames of detection data. By simply extracting one frame of said detection data every certain frames, the purpose of sampling said input signals at a preset time interval can be realized. By simply de-interleaving the detection data extracted, the input signals of every one of audio input interfaces can be re-acquired, which is very convenient.
Furthermore, after the input signals of every one of said audio input interfaces have been acquired, each of the input signals is pre-processed so as to ensure the accuracy of the result of the detection. The pre-processing comprises high-pass filtering, filtering certain frequency interferences, noise suppression, and so on, so as to reduce the influence of the noise on the detection of the input signals.
With respect to S102, the energy values of the input signals of every one of said audio input interfaces acquired herein are detect ed and acquired, each energy value represents the intensity of a respective input signal. The energy value of an input signal is the maximum among the energy values, means that the intensity of the input signal is the strongest, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user. The interface identification of said audio input interface is acquired to identify the audio input interface of the input signal whose intensity is the strongest in the present detection.
With respect to S103, said interface identification acquired herein is added into an identification sequence in an order by acquiring time. Said identification sequence may be created in a buffer or in other types of memories, so that it is easy to be accessed.
Preferably, said identification sequence follows the rule of first-in-and-first-out in an order by acquiring time, and the quantity of the interface identifications saved each time is less than or equal to a preset quantity. That is, if the quantity of interface identifications saved in said identification sequence reaches said preset quantity, whenever a new interface identification is added, an interface identification of which the acquiring time (in another word, the save time) is the earliest will be discarded. The quantity of theinterfaceidentifications saved in saididentification sequence is maintained in said preset quantity, such as 25, and the most recently acquired 25 of interface identifications will always be saved. The interface identifications saved in the identification sequence may be the same, that is, said identification sequence may contain multiple of the same interface identification.
S104: identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
The more often the same interface identifications is saved in said identification sequence, means that the more often the audio frequency with the maximum value is input into the audio input interface corresponding to said interface identification, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user, and the correct chance will be high to identify said audio input interface corresponding to said interface identification as a valid audio input interface. And with the increasing of the quantity of the interface identifications saved in said identification sequence, the accuracy of identifying will be further increased. Moreover, after the valid audio input interface has been identified, said audio input interface may be automatically matched to an audio software in the background for processing. Alternatively, the input signal of said audio input interface may be subject to processing such as filtering and so on, before it is output into said audio software in the background for processing. After said audio input interface has been identified, a user interface may be further displayed, on which the audio input interfaceidentified currently or the microphone or other audio input devices connected to said audio input interface will be shown. The start and the end of the method for detecting an audio input interface of the present disclosure may be triggered through many different modes, for example, it may be set to start the detection when it is detected that the microphone is inserted, or it may be set to start the detection when a start instruction is received; also for example, it may be set to end the detection when the right input interface is detected, or it may be set to end the detection when the microphone is removed, or it may be set to end the detection when an end instruction is received, and so on.
As shown in Fig. 2, which is a schematic flow diagram illustrating the method for detecting an audio input interface according to the second embodiment of the present disclosure, the step S102 is specified as following sub-steps:
S1022 : acquiring energy values of the input signals of every one of said audio input interfaces ;
S1024: judging whether said maximum energy value among the energy values of every one of said input signals is no less than (greater than or equal to) a preset energy value;
if it is, then going to S1026; if it is not, then judging the input signals of every one of the audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces;
S1026: acquiring the interface identification of the audio input interface of which the energy value is the maximum among the energy values.
The method for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 1 in that, prior to said acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values (S1026), the method further comprises step S1024.
According to the method for detecting an audio input interface in this embodiment, the interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired only when it is judged that said maximum energy value among the energy values of every one of said input signals is greater than or equal to a preset energy value, otherwise, the input signals of every one of said audio input interfaces are judged as all invalid. As a result, if the audio input with the maximum energy value is created by a noise, said audio input will be judged as an invalid signal as long as the intensity of noise is less than said preset energy value, and theresult of the identificationwill not be affected, the influence of the noise on the result of the identification is effectively reduced.
As shown in Fig. 3, which is a schematic flow diagram illustrating the method for detecting an audio input interface according to the third embodiment of the present disclosure, the method for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 1 in that, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step S105:
S105: executing VAD detection on the input signals of every one of said audio input interfaces;
if all results of said VAD detection on every one of said audio input interfaces are zero (0) , then judging said input signals of every one of said audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces ;
if at least one of the results of said VAD detection of every one of said audio input interfaces is one (1), then judging the input signals of every one of audio input interfaces as valid, and going to step S102 to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
Said VAD detection is so called voice activity detection, which can effectively detect the activity of the input signals, identify the input signal which may be the audio input, and increase the speed of identifying the activities audio input interface. If all results of the VAD detection are 0, that means every one of said audio input interfaces are currently in a muted state; if at least one of results of said audio input interfaces is 1, then at least one of said audio input interfaces has an audio input, and the input signals of every one of audio input interfaces collected herein can be judged as valid, and then the next step (acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values) is executed.
As shown in Fig. 4, which is a schematic flow diagram illustrating the method for detecting an audio input interface according to the fourth embodiment of the present disclosure, the method for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 1 in that, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step S106:
acquiring signal-noise ratios of the input signals of every one of said audio input interfaces ;
if each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then judging the input signals of every one of said audio input interfaces as invalid, and re-acquiring input signals of every one of said audio input interfaces;
if at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then judging every one of said input signals as valid, and going to step S102 to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
The so called signal-noise ratio is a ratio of a normal sound signal to a noise signal (power) with no-sound signal, and it is often represented in the unit of dB. In this embodiment, the signal-noise ratios of every one of audio input signals are detected, and the input signals of the audio input interface are judged as valid only when at least one of the signal-noise ratios of said audio input interfaces is no less than the preset value of signal-noise ratio, otherwise, the input signals of the audio input interface are judged as invalid. As a result, the influence of the noise on identifying an active audio input interface is reduced, and the accuracy of identifying is improved.
In a preferred embodiment, any two or any three of steps S1021, S105 and S106 can be selected and executed in combination simultaneously, thereby further improving the accuracy and efficiency of identifying.
In another preferred embodiment of the method, after said identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface, further executing a step of post-processing to the audio input interface identified by adjusting parameters related to the device so as to make the microphone connected to said device in the best working state.
Said post-processing comprises: executing a self-adaptive microphone volume adjustment to said valid audio input interface, that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state; executing a signal-noise ratio detection, starting a noise suppression according to the result of the detection, and so on.
According to the method of the present disclosure, the user does not need to participate in setting the device, but the device will automatically select the microphone, to which the user is inputting voice signals, and the problem of no sound will never happen, and the corresponding result of the configuration will be displayed in a user interface. When there are multiple microphones with different acoustic characteristics connected to one device, the user simply needs to speak to the microphone he wants, then the automatically switching of the audio input interfaces will be achieved, without the need of manual setting. If the microphone is broken and no sound is collected, the device can also automatically switch off the corresponding audio input interface, without the need of manual setting. With the method of the present disclosure, the user who doesn't know how to set the microphone is relieved from the difficulty, which is a convenience for the user.
As shown in Fig. 5, which is a structure diagram illustrating the system for detecting an audio input interface according to the first embodiment of the present disclosure, said system for detecting an audio input interface of present disclosure comprises:
an input detecting module 11, configured to acquire input signals of every one of audio input interfaces;
an energy detecting module 12, configured to acquire signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
an interface identification acquiring module 13, configured to add said interface identification acquired herein into an identification sequence in an order by acquiring time;
an identifying module 14, configured to the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
According to the system for detecting an audio input interface of the present disclosure, input signals of every one of audio input interfaces are acquired at a preset time interval, energy values of said input signals of every one of said audio input interfaces are detected, and an interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired. Each of the interface identifications acquired is added to a preset identification sequence in an order by acquiring time. If the same interface identification is saved more often in said identification sequence, it means that, there are more such situations that the energy value of the input signal of the audio input interface corresponding to said interface identification is the maximum, and the audio input interface corresponding to said interface identification will be identified as a valid audio input interface. The system of the present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, and the user does not need to switch manually, which is very convenient, moreover, the influence of the noise on the result of identifying is reduced, thereby the result of identifying is more accurate.
Wherein, with respect to said input detecting module 11, the input signals of every one of audio input interfaces can be acquired through monitoring for and collecting data from each of said audio input interfaces. The input signals of every one of said audio input interfaces comprises input signals from a microphone hardware-connected to any one of said audio input interfaces, noise signals, and so on. In one embodiment, all of the audio input interfaces can be enumerated by means of calling a function of Dsound API (Direct Sound Capture Enumerate( )).
After all of the audio input interfaces of the device having been detected, the input data of every one of the audio input interfaces is acquired by means of collecting the input signals of every one of said audio input interfaces. Preferably, the parameter of each of the audio input interfaces is preset to unify the audio collecting format for each of the audio input interfaces, such as using an audio collecting format with mono-channel, and 44.1 KHz of sampling rate. By unifying the audio collecting format of each of the audio input interfaces, a large amount of computational load can be reduced in the post processing of the input signals, and the speed of identifying the microphone can be increased.
In one embodiment, said input detecting module 11 comprises following sub-modules:
a collecting unit, configured to simultaneously collecting input signals of every one of said audio input interfaces ;
an encapsulating unit, configured to encapsulate said input signals of every one of said audio input interfaces being collected at the same time into a frame of detection data;
an extracting unit, configured to de-interleave each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces contained in said frame of detection data.
Moreover, said input detecting module 11 further comprises a saving unit, configured to save each of frames of said detection data; said extracting unit is further configured to extract one of said frames of detection data every preset frames, and then de-interleave said frame extracted.
For example, when the encapsulating unit encapsulates the input signals of every one of said audio input interfaces, each of the input signals is collected in an unit of 20 milliseconds, and then signals of every one of the input signals are put in a buffer with a M*20 ( M represents the number of the devices enumerated ) milliseconds length, the corresponding data is acquired from said buffer and encapsulated.
In this way, the input signals (such as N channels of input signals, N is a nature number) of every one of audio input interfaces are encapsulated in respective frames of detection data. By simply extracting one frame of said detection data every certain frames, the purpose of sampling said input signals at a preset time interval can be realized. By simply de-interleaving the detection data extracted, the input signals of every one of audio input interfaces can be re-acquired, which is very convenient.
Furthermore, after the input detecting module 11 has acquired the input signals of every one of said audio input interfaces, each of the input signals is pre-processed so as to ensure the accuracy of the result of the detection. The pre-processing comprises high-pass filtering, filtering certain frequency interferences, noise suppression, and so on, so as to reduce the influence of the noise on the detection of the input signals.
With respect to said energy detecting module 12, the energy values of the input signals of every one of said audio input interfaces acquired herein are detected and acquired, each energy value represents the intensity of a respective input signal. The energy value of an input signal is the maximum among the energy values, means that the intensity of the input signal is the strongest, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user. The interface identification of said audio input interface is acquired to identify the audio input interface of the input signal whose intensity is the strongest in the present detection.
With respect to said interface identification acquiring module 13, said interface identification acquired herein is added into an identification sequence in an order by acquiring time. Said identification sequence may be created in a buffer or in other types of memories, so that it is easy to be accessed.
Preferably, said identification sequence follows the rule of first-in-and-first-out in an order by acquiring time, and the quantity of the interface identifications saved each time is less than or equal to a preset quantity. That is, if the quantity of interface identifications saved in said identification sequence reaches said preset quantity, whenever a new interface identification is added, an interface identification of which the acquiring time (in another word, the save time) is the earliest will be discarded. The quantity of the interface identifications saved in said identification sequence is maintained in said preset quantity, such as 25, and the most recently acquired 25 of interface identifications will always be saved. The interface identifications saved in the identification sequence may be the same, that is, said identification sequence may contain multiple of the same interface identification.
With respect to said identifying module 14, it is configured to identify the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
The more often the same interface identifications is saved in said identification sequence, means that the more often the audio frequency with the maximum value is input into the audio input interface corresponding to said interface identification, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user, and the correct chance will be high to identify said audio input interface corresponding to said interface identification as a valid audio input interface. And with the increasing of the quantity of the interface identifications saved in said identification sequence, the accuracy of identifying will be further increased.
Moreover, after the valid audio input interface has been identified, said audio input interface may be automatically matched to an audio software in the background for processing. Alternatively, the input signal of said audio input interface may be subject to processing such as filtering and so on, before it is output into said audio software in the background for processing. After said audio input interface has been identified, a user interface may be further displayed, on which the audio input interface identified currently or the microphone or other audio input devices connected to said audio input interface will be shown.
The start and the end of the system for detecting an audio input interface of the present disclosure may be triggered through many different modes, for example, it may be set to start the detection when it is detected that the microphone is inserted, or it may be set to start the detection when a start instruction is received; also for example, it may be set to end the detection when the right input interface is detected, or it may be set to end the detection when the microphone is removed, or it may be set to end the detection when an end instruction is received, and so on.
The present disclosure further provides a second embodiment of the system for detecting an audio input interface, which is mainly different from the first embodiment as shown in Fig. 5 in that, prior to acquiring an interface identification of said audio input interface of which the energy value is the maximum among the energy values, said interface identification acquiring module 13 further judges whether said maximum energy value among the energy values of every one of said input signals is no less than (greater than or equal to) a preset energy value; if it is, then acquires the interface identification of the audio input interface of which the energy value is the maximum among the energy values; if it is not, then judges the input signals of every one of the audio input interfaces as all invalid, and re-acquires input signals of every one of said input interfaces;
According to the system for detecting an audio input interface in this embodiment, said interface identification acquiring module 13 acquires the interface identification of the audio input interface of which the energy value is the maximum among the energy values only when it is judged that said maximum energy value among the energy values of every one of said input signals is greater than or equal to a preset energy value, otherwise, the input signals of every one of said audio input interfaces are judged as all invalid. As a result, if the audio input with the maximum energy value is created by a noise, said audio input will be judged as an invalid signal as long as the intensity of noise is less than said preset energy value, and the result of the identification will not be affected, the influence of the noise on the result of the identification is effectively reduced.
As shown in Fig. 6, which is a structure diagram illustrating the system for detecting an audio input interface according to the third embodiment of the present disclosure, the system for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 5 in that, the system further comprises a VAD detecting module 15;
Said VAD detecting module 15 is configured to execute VAD detection on the input signals of every one of said audio input interfaces;
if all results of said VAD detection on every one of said audio input interfaces are zero (0), then judging said input signals of every one of said audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces; if at least one of the results of said VAD detection of every one of said audio input interfaces is one (1), then judging the input signals of every one of audio input interfaces as valid, and said energy detecting module 12 acquires energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
Said VAD detection is so called voice activity detection, which can effectively detect the activity of the input signals, identify the input signal which may be the audio input, and increase the speed of identifying the activities audio input interface. If all results of the VAD detection are 0, that means every one of said audio input interfaces are currently in a muted state; if at least one of results of said audio input interfaces is 1, then at least one of said audio input interfaces has an audio input, and the input signals of every one of audio input interfaces collected herein can be judged as valid, and then the next step (acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values) is executed.
As shown in F ig . 7, which is a structure diagram illustrating the system for detecting an audio input interface according to the fourth embodiment of the present disclosure, the system is mainly different from the first embodiment as shown in Fig. 5 in that, the system further comprises a signal-noise ratio detecting module 16, which is configured to acquire signal-noise ratios of the input signals of every one of said audio input interfaces; if each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then judging the input signals of every one of said audio input interfaces as invalid, and re-acquiring input signals of every one of said audio input interfaces; if at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then judging every one of said input signals as valid, said energy detecting module 12 acquires energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
The so called signal-noise ratio is a ratio of a normal sound signal to a noise signal (power) with no-sound signal, and it is often represented in the unit of dB. In this embodiment, the signal-noise ratios of every one of audio input signals are detected, and the input signals of the audio input interface are judged as valid only when at least one of the signal-noise ratios of said audio input interfaces is no less than the preset value of signal-noise ratio, otherwise, the input signals of the audio input interface are judged as invalid. As a result, the influence of the noise on identifying an active audio input interface is reduced, and the accuracy of identifying is improved.
In a preferred embodiment, any two or any three of said interface identification acquiring module 13, said VAD detecting module 15 and said signal-noise ratio detecting module 16 can be selected and adopted in combination, thereby further improving the accuracy and efficiency of identifying.
In another preferred embodiment of the system, after said identifying module 14 identifies the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface, post-processing is further executed to the audio input interface identified by adjusting parameters related to the device so as to make the microphone connected to said device in the best working state.
Said post-processing comprises: executing a self-adaptive microphone volume adjustment to said valid audio input interface, that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state; executing a signal-noise ratio detection, starting a noise suppression according to the result of the detection, and so on.
According to the system of the present disclosure, the user does not need to participate in setting the device, but the device will automatically select the microphone, to which the user is inputting voice signals, and the problem of no sound will never happen, and the corresponding result of the configuration will be displayed in a user interface. When there are multiple microphones with different acoustic characteristics connected to one device, the user simply needs to speak to the microphone he wants, then the automatically switching of the audio input interfaces will be achieved, without the need of manual setting. If the microphone is broken and no sound is collected, the device can also automatically switch off the corresponding audio input interface, without the need of the user's manual setting. With the system of the present disclosure, the user who doesn't know how to set the microphone is relieved from the difficulty, which is a convenience for the user.
The person skilled in the art can understand that, all of or part of the processes and the corresponding system implementing the embodiments mentioned above, may be achieved by means of relevant hardware commanded by computer programs, the computer programs may be saved in the computer readable storage medium, and they may comprise the processes of embodiments of the respective methods and systems mentioned above when the programs are executed. Wherein, the storage medium may be a disk or CD or read-only memory or random access memory, and so on.
The foregoing examples are preferred embodiments of the present disclosure only and not intended to limit the present disclosure. It should be understood that, to the person skilled in the art, various modifications and improvements can be made without departing from the spirit and principle of the present disclosure, which should all be included within the scope of the present disclosure. Therefore, the protection scope of the present disclosure shall be defined by the appended claims.

Claims (20)

  1. A methodfor detecting an audio input interface , comprising :
    acquiring input signals of every one of audio input interfaces;
    acquiring energy values of said input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
    adding said interface identification acquired herein into an identification sequence in an order by acquiring time ; and
    identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
  2. The method for detecting an audio input interface according to claim 1, further comprising:
    when the quantity of interface identifications saved in said identification sequence reaches a preset quantity, whenever a new interface identification is added, one saved interface identification of which the acquiring time is the earliest will be discarded.
  3. The method for detecting an audio input interface according to claim 1 or 2, wherein, prior to said acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values, the method further comprises the following sub-step:
    judging whether said maximum energy value among the energy values of every one of said input signals is no less than a preset energy value;
    when said maximum energy value among the energy values of every one of said input signals is no less than the preset energy value, then acquiring the interface identification of the audio input interface of which the energy value is the maximum among the energy values; when said maximum energy value among the energy values of every one of said input signals is less than the preset energy value , then judging the input signals of every one of the audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces.
  4. The method for detecting an audio input interface according to claim 1 or 2, wherein, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step:
    executing VAD detection on the input signals of every one of said audio input interfaces;
    when all results of said VAD detection on every one of said audio input interfaces are zero, then judging said input signals of every one of said audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces;
    when at least one of the results of said VAD detection of every one of said audio input interfaces is one, then judging the input signals of every one of audio input interfaces as valid, acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
  5. The method for detecting an audio input interface according to claim 1 or 2, wherein, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step:
    acquiring signal-noise ratios of the input signals of every one of said audio input interfaces;
    when each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then judging the input signals of every one of said audio input interfaces as invalid, and re-acquiring input signals of every one of said audio input interfaces;
    when at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then judging every one of said input signals as valid, acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
  6. The method for detecting an audio input interface according to claim 1 or 2, wherein, said acquiring input signals of every one of audio input interfaces comprises following sub-steps:
    simultaneously collecting input signals of every one of said audio input interfaces, and encapsulating said input signals of every one of said audio input interfaces collected at the same time into a frame of detection data;
    de-interleaving each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces.
  7. The method for detecting an audio input interface according to claim 6, wherein, said de-interleaving each of frames of said detection data comprises the following sub-steps:
    saving each of frames of said detection data ;
    extracting one of said frames of detection data every preset frames, and then de-interleaving said frame extracted herein.
  8. A system for detecting an audio input interface , comprising:
    an input detecting module, configured to acquire input signals of every one of audio input interfaces;
    an energy detecting module, configured to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
    an interface identification acquiring module, configured to add said interface identification acquired herein into an identification sequence in an order by acquiring time; and
    an identifying module, configured to identify the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
  9. The system for detecting an audio input interface according to claim 8 , wherein, the interface identification is configured to discard one saved interface identification of which the acquiring time is the earliest whenever a new interface identification is added, if the quantity of interface identifications saved in said identification sequence reaches a preset quantity.
  10. The system for detecting an audio input interface according to claim 8 or 9, wherein, said interface identification acquiring module is configured to judge whether said maximum energy value among the energy values of every one of said input signals is no less than a preset energy value;
    when said maximum energy value among the energy values of every one of said input signals is no less than a preset energy value, then the system acquires the interface identification of the audio input interface of which the energy value is the maximum among the energy values; when said maximum energy value among the energy values of every one of said input signals is less than a preset energy value, then the system judges the input signals of every one of the audio input interfaces as all invalid, and re-acquires input signals of every one of said input interfaces.
  11. The system for detecting an audio input interface according to claim 8 or 9, furthering comprising:
    a VAD detecting module, configured to execute VAD detection on the input signals of every one of said audio input interfaces;
    when all results of said VAD detection on every one of said audio input interfaces are zero, then the system judges said input signals of every one of said audio input interfaces as all invalid, and re-acquires input signals of every one of said input interfaces;
    when at least one of the results of said VAD detection of every one of said audio input interfaces is one, then the system judges the input signals of every one of audio input interfaces as valid, the energy detecting module acquires energy values of the input signals of every one of said audio input interfaces, and acquires an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
  12. The system for detecting an audio input interface according to claim 8 or 9, furthering comprising :
    a signal-noise ratio detecting module, configured to acquire signal-noise ratios of the input signals of every one of said audio input interfaces;
    when each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then the system judges the input signals of every one of said audio input interfaces as invalid, and re-acquires input signals of every one of said audio input interfaces;
    when at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then the system judges every one of said input signals as valid, the energy detecting module acquires energy values of the input signals of every one of said audio input interfaces, and acquires an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
  13. The system for detecting an audio input interface according to claim 8 or 9, wherein, said input detecting module comprises following sub-modules:
    a collecting unit, configured to simultaneously collecting input signals of every one of said audio input interfaces ;
    an encapsulating unit, configured to encapsulate said input signals of every one of said audio input interfaces being collected at the same time into a frame of detection data;
    an extracting unit, configured to de-interleave each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces contained in said frame of detection data.
  14. The system for detecting an audio input interface according to claim 13, wherein, said input detecting module further comprises a saving unit, which is configured to save each of frames of said detection data; and wherein said extracting unit is further configured to extract one of said frames of detection data every preset frames, and then de-interleave said frame extracted.
  15. One or more non-transitory computer readable storage medium, including computer executable instructions, aid executable instructions are configured to execute a method for detecting an audio input interface, wherein, said method comprising:
    acquiring input signals of every one of audio input interfaces;
    acquiring energy values of said input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
    adding said interface identification acquired herein into an identification sequence in an order by acquiring time ; and
    identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
  16. The one or more non-transitory computer readable storage medium according to claim 15, wherein, said method further comprising:
    when the quantity of interface identifications saved in said identification sequence reaches a preset quantity, whenever a new interface identification is added, one saved interface identification of which the acquiring time is the earliest will be discarded.
  17. The one or more non-transitory computer readable storage medium according to claim 16, wherein, prior to said acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values, the method further comprises the following sub-step:
    judging whether said maximum energy value among the energy values of every one of said input signals is no less than a preset energy value;
    when said maximum energy value among the energy values of every one of said input signals is no less than the preset energy value, then acquiring the interface identification of the audio input interface of which the energy value is the maximum among the energy values; when said maximum energy value among the energy values of every one of said input signals is less than the preset energy value , then judging the input signals of every one of the audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces.
  18. The one or more non-transitory computer readable storage medium according to claim 16, wherein, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step:
    executing VAD detection on the input signals of every one of said audio input interfaces;
    when all results of said VAD detection on every one of said audio input interfaces are zero, then judging said input signals of every one of said audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces;
    when at least one of the results of said VAD detection of every one of said audio input interfaces is one, then judging the input signals of every one of audio input interfaces as valid, acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
  19. The one or more non-transitory computer readable storage medium according to claim 16, wherein, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step:
    acquiring signal-noise ratios of the input signals of every one of said audio input interfaces;
    when each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then judging the input signals of every one of said audio input interfaces as invalid, and re-acquiring input signals of every one of said audio input interfaces;
    when at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then judging every one of said input signals as valid, acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
  20. The one or more non-transitory computer readable storage medium according to claim 16, wherein, said acquiring input signals of every one of audio input interfaces comprises following sub-steps:
    simultaneously collecting input signals of every one of said audio input interfaces, and encapsulating said input signals of every one of said audio input interfaces collected at the same time into a frame of detection data;
    de-interleaving each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces.
PCT/CN2014/075663 2013-05-27 2014-04-18 Method, system and computer storage medium for detecting an audio input interface Ceased WO2014190824A1 (en)

Priority Applications (1)

Application Number Priority Date Filing Date Title
US14/683,103 US9886238B2 (en) 2013-05-27 2015-04-09 Method, system and computer storage medium for detecting an audio input interface

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
CN201310202043.3A CN103327433B (en) 2013-05-27 2013-05-27 Audio input interface detection method and system thereof
CN201310202043.3 2013-05-27

Related Child Applications (1)

Application Number Title Priority Date Filing Date
US14/683,103 Continuation US9886238B2 (en) 2013-05-27 2015-04-09 Method, system and computer storage medium for detecting an audio input interface

Publications (1)

Publication Number Publication Date
WO2014190824A1 true WO2014190824A1 (en) 2014-12-04

Family

ID=49195917

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/CN2014/075663 Ceased WO2014190824A1 (en) 2013-05-27 2014-04-18 Method, system and computer storage medium for detecting an audio input interface

Country Status (3)

Country Link
US (1) US9886238B2 (en)
CN (1) CN103327433B (en)
WO (1) WO2014190824A1 (en)

Families Citing this family (8)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
CN103327433B (en) * 2013-05-27 2014-08-27 腾讯科技(深圳)有限公司 Audio input interface detection method and system thereof
WO2016015186A1 (en) * 2014-07-28 2016-02-04 华为技术有限公司 Acoustical signal processing method and device of communication device
CN105788609B (en) * 2014-12-25 2019-08-09 福建凯米网络科技有限公司 The correlating method and device and assessment method and system of multichannel source of sound
CN105681974A (en) * 2016-04-01 2016-06-15 北京小鸟听听科技有限公司 Multiple-sound-source switching method and device and audio equipment
CN111354356B (en) * 2018-12-24 2024-04-30 北京搜狗科技发展有限公司 A method and device for processing voice data
CN111736795B (en) * 2019-06-24 2024-11-29 北京京东尚科信息技术有限公司 Audio processing method, device, equipment and storage medium
CN111770427B (en) * 2020-06-24 2023-01-24 杭州海康威视数字技术股份有限公司 Microphone array detection method, device, equipment and storage medium
CN113542975A (en) * 2021-07-02 2021-10-22 南昌华勤电子科技有限公司 Audio signal switching circuit and electronic equipment

Citations (4)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
CN1805489A (en) * 2005-01-10 2006-07-19 华为技术有限公司 Method and system of implementing report of current speaker during conference
US20080040117A1 (en) * 2004-05-14 2008-02-14 Shuian Yu Method And Apparatus Of Audio Switching
CN102056053A (en) * 2010-12-17 2011-05-11 中兴通讯股份有限公司 Multi-microphone audio mixing method and device
CN103327433A (en) * 2013-05-27 2013-09-25 腾讯科技(深圳)有限公司 Audio input interface detection method and system thereof

Family Cites Families (7)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US6956828B2 (en) * 2000-12-29 2005-10-18 Nortel Networks Limited Apparatus and method for packet-based media communications
CN1271593C (en) * 2004-12-24 2006-08-23 北京中星微电子有限公司 Voice signal detection method
CN1845580A (en) * 2005-04-07 2006-10-11 深圳Tcl新技术有限公司 Audio and video signal source recognition and automatic switching method and apparatus
CN101308651B (en) * 2007-05-17 2011-05-04 展讯通信(上海)有限公司 Detection method of audio transient signal
CN101299782B (en) * 2008-05-22 2011-09-21 杭州华三通信技术有限公司 Method and device for detecting double-audio signal
US8514265B2 (en) * 2008-10-02 2013-08-20 Lifesize Communications, Inc. Systems and methods for selecting videoconferencing endpoints for display in a composite video image
CN102143262B (en) * 2010-02-03 2014-03-26 深圳富泰宏精密工业有限公司 Electronic device and method for switching audio input channel thereof

Patent Citations (4)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US20080040117A1 (en) * 2004-05-14 2008-02-14 Shuian Yu Method And Apparatus Of Audio Switching
CN1805489A (en) * 2005-01-10 2006-07-19 华为技术有限公司 Method and system of implementing report of current speaker during conference
CN102056053A (en) * 2010-12-17 2011-05-11 中兴通讯股份有限公司 Multi-microphone audio mixing method and device
CN103327433A (en) * 2013-05-27 2013-09-25 腾讯科技(深圳)有限公司 Audio input interface detection method and system thereof

Also Published As

Publication number Publication date
CN103327433B (en) 2014-08-27
US9886238B2 (en) 2018-02-06
CN103327433A (en) 2013-09-25
US20150212792A1 (en) 2015-07-30

Similar Documents

Publication Publication Date Title
WO2014190824A1 (en) Method, system and computer storage medium for detecting an audio input interface
US8233637B2 (en) Multi-membrane microphone for high-amplitude audio capture
US7968786B2 (en) Volume adjusting apparatus and volume adjusting method
US9020157B2 (en) Active noise cancellation system
KR101118217B1 (en) Audio data processing apparatus and method therefor
US6549630B1 (en) Signal expander with discrimination between close and distant acoustic source
EP2731351A2 (en) Adaptive system for managing a plurality of microphones and speakers
WO2020060206A1 (en) Methods for audio processing, apparatus, electronic device and computer readable storage medium
EP2426950A2 (en) Noise suppression for sending voice with binaural microphones
WO2019037319A1 (en) Electric quantity early warning method, and server, mobile terminal and storage medium
JP2012105287A (en) Intelligibility control using ambient noise detection
WO2019051890A1 (en) Terminal control method and device, and computer-readable storage medium
US9672843B2 (en) Apparatus and method for improving an audio signal in the spectral domain
JP2003169119A (en) Speaker assembly for mobile phone
WO2018139884A1 (en) Method for processing vr audio and corresponding equipment
WO2014148844A1 (en) Terminal device and audio signal output method thereof
KR20140055932A (en) Apparatus and method for keeping output loudness and quality of sound among differernt equalizer modes
EP3748635B1 (en) Acoustic device and acoustic processing method
WO2014148845A1 (en) Audio signal size control method and device
WO2019084886A1 (en) Audio denoising method and denoising device, mobile terminal and readable storage medium
CN113286244A (en) Microphone anomaly detection method and device
WO2019019329A1 (en) Method and system for optimizing network performance and computer readable storage medium
GB2500251A (en) Active noise cancellation system with wind noise reduction
US9124985B2 (en) Hearing aid and method for automatically controlling directivity
WO2015180430A1 (en) Voice control method and system

Legal Events

Date Code Title Description
121 Ep: the epo has been informed by wipo that ep was designated in this application

Ref document number: 14804561

Country of ref document: EP

Kind code of ref document: A1

NENP Non-entry into the national phase

Ref country code: DE

32PN Ep: public notification in the ep bulletin as address of the adressee cannot be established

Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205 DATED 07/04/2016)

122 Ep: pct application non-entry in european phase

Ref document number: 14804561

Country of ref document: EP

Kind code of ref document: A1