WO2014190824A1 - Method, system and computer storage medium for detecting an audio input interface - Google Patents
Method, system and computer storage medium for detecting an audio input interface Download PDFInfo
- Publication number
- WO2014190824A1 WO2014190824A1 PCT/CN2014/075663 CN2014075663W WO2014190824A1 WO 2014190824 A1 WO2014190824 A1 WO 2014190824A1 CN 2014075663 W CN2014075663 W CN 2014075663W WO 2014190824 A1 WO2014190824 A1 WO 2014190824A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- audio input
- interface
- input signals
- interfaces
- acquiring
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F3/00—Input arrangements for transferring data to be processed into a form capable of being handled by the computer; Output arrangements for transferring data from processing unit to output unit, e.g. interface arrangements
- G06F3/16—Sound input; Sound output
- G06F3/167—Audio in a user interface, e.g. using voice commands for navigating, audio feedback
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04R—LOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
- H04R3/00—Circuits for transducers
- H04R3/005—Circuits for transducers for combining the signals of two or more microphones
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
- G10L25/78—Detection of presence or absence of voice signals
- G10L2025/783—Detection of presence or absence of voice signals based on threshold decision
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04M—TELEPHONIC COMMUNICATION
- H04M3/00—Automatic or semi-automatic exchanges
- H04M3/42—Systems providing special services or facilities to subscribers
- H04M3/56—Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities
- H04M3/561—Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities by multiplexing
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04M—TELEPHONIC COMMUNICATION
- H04M3/00—Automatic or semi-automatic exchanges
- H04M3/42—Systems providing special services or facilities to subscribers
- H04M3/56—Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities
- H04M3/568—Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities audio processing specific to telephonic conferencing, e.g. spatial distribution, mixing of participants
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04M—TELEPHONIC COMMUNICATION
- H04M3/00—Automatic or semi-automatic exchanges
- H04M3/42—Systems providing special services or facilities to subscribers
- H04M3/56—Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities
- H04M3/568—Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities audio processing specific to telephonic conferencing, e.g. spatial distribution, mixing of participants
- H04M3/569—Arrangements for connecting several subscribers to a common circuit, i.e. affording conference facilities audio processing specific to telephonic conferencing, e.g. spatial distribution, mixing of participants using the instant speaker's algorithm
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04R—LOUDSPEAKERS, MICROPHONES, GRAMOPHONE PICK-UPS OR LIKE ACOUSTIC ELECTROMECHANICAL TRANSDUCERS; ELECTRIC HEARING AIDS; PUBLIC ADDRESS SYSTEMS
- H04R27/00—Public address systems
Definitions
- the present disclosure relates to the field of audio processing , and more particularly to a method and a device for detecting an audio input interface.
- an object of the present disclosure is to provide a method for detecting an audio input interface, by means of detecting the input data of every one of audio input interfaces, the audio input interface connected to the microphone, into which a voice signal is being input, can be effectively identified, which does not need the user to switch manually and is very convenient.
- the present disclosure is realized by the following technical scheme:
- a method for detecting an audio input interface comprises:
- the present disclosure is to provide a system for detecting an audio input interface, comprising:
- an input detecting module configured to acquire input signals of every one of audio input interfaces
- an energy detecting module configured to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
- an interface identification acquiring module configured to add said interface identification acquired herein into an identification sequence in an order by acquiring time
- an identifying module configured to identify the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
- input signals of every one of audio input interfaces are acquired at a preset time interval, energy values of said input signals of every one of said audio input interfaces are detected, and an interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired.
- Each of the interface identifications acquired is added to a preset identification sequence in an order by acquiring time. If the same interface identification is saved more often in said identification sequence, it means that, there are more such situations that the energy value of the input signal of the audio input interface corresponding to said i nterface identification is the maximum, and the audio input interface corresponding to said interface identification will be identified as a valid audio input interface.
- the present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, and the user does not need to switch manually, which is very convenient, moreover, the influence of the noise on the result of identifying is reduced, thereby the result of identifying is more accurate .
- FIG. 1 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the first embodiment of the present disclosure ;
- FIG. 2 is a schematic flow diagram illustrating the method for detecting an audio input interface according tothe second embodiment of the present disclosure
- FIG. 3 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the third embodiment of the present disclosure ;
- FIG. 4 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the fourth embodiment of the present disclosure
- FIG. 5 is a structure diagram illustrating the system for detecting an audio input interface according to the first embodiment of the present disclosure
- FIG. 6 is a structure diagram illustrating the system for detecting an audio input interface according to the third embodiment of the present disclosure.
- FIG. 7 is a structure diagram illustrating the system for detecting an audio input interface according to the fourth embodiment of the present disclosure.
- Fig. 1 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the first embodiment of the present disclosure
- the method for detecting an audio input interface of said embodiment comprises following steps:
- S101 acquiring input signals of every one of audio input interfaces ;
- S102 acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
- S104 identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
- input signals of every one of audio input interfaces are acquired at a preset time interval, energy values of said input signals of every one of said audio input interfaces are detected, and an interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired.
- Each of the interface identifications acquired is added to a preset identification sequence in an order by acquiring time. If the same interface identification is saved more often in said identification sequence, it means that, there are more such situations that the energy value of the input signal of the audio input interface corresponding to said interface identification is the maximum, and the audio input interface corresponding to said interface identification will be identified as a valid audio input interface.
- the method of the present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, and the user does not need to switch manually, which is very convenient, moreover, the influence of the noise on the result of identifying is reduced, thereby the result of identifying is more accurate.
- the input signals of every one of audio input interfaces can be acquired through monitoring for and collecting data from each of saidaudio input interfaces.
- the input signals of every one of said audio input interfaces comprise input signals from a microphone hardware-connected to any one of said audio input interfaces, noise signals, and so on.
- all of the audio input interfaces can be enumerated by means of calling a function of Dsound API (Direct Sound Capture Enumerate( )).
- the input data of every one of theaudio input interfaces is acquired by means of collecting the input signals of every one of said audio input interfaces.
- the parameter of each of the audio input interfaces is preset to unify the audio collecting format for each of the audio input interfaces, such as using an audio collecting format with mono-channel, and 44.1 KHz of sampling rate.
- input signals of every one of audio input interfaces are acquired simultaneously at a preset time interval, which comprises following sub-steps:
- S1011 simultaneously collecting input signals of every one of said audio input interfaces, and encapsulating said input signals of every one of said audio input interfaces collected at the same time into a frame of detection data.
- S1012 de-interleaving each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces. Moreover, each of frames of said detection data can be saved; one of said frames of detection data is extracted every preset frames and further de-interleaved so as to acquire the input signals of every one of said audio input interfaces contained in said frame of detection data .
- each of the input signals is collected in an unit of 20 milliseconds, and then signals of every one of the input signals are put in a buffer with a M*20(M represents the number of the devices enumerated)milliseconds length, the corresponding data is acquired from said buffer and encapsulated.
- the input signals (such as N channels of input signals, N is a nature number) of every one of audio input interfaces are encapsulated in respective frames of detection data.
- the purpose of sampling said input signals at a preset time interval can be realized.
- the input signals of every one of audio input interfaces can be re-acquired, which is very convenient.
- each of the input signals is pre-processed so as to ensure the accuracy of the result of the detection.
- the pre-processing comprises high-pass filtering, filtering certain frequency interferences, noise suppression, and so on, so as to reduce the influence of the noise on the detection of the input signals.
- each energy value represents the intensity of a respective input signal.
- the energy value of an input signal is the maximum among the energy values, means that the intensity of the input signal is the strongest, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user.
- the interface identification of said audio input interface is acquired to identify the audio input interface of the input signal whose intensity is the strongest in the present detection.
- said interface identification acquired herein is added into an identification sequence in an order by acquiring time.
- Said identification sequence may be created in a buffer or in other types of memories, so that it is easy to be accessed.
- said identification sequence follows the rule of first-in-and-first-out in an order by acquiring time, and the quantity of the interface identifications saved each time is less than or equal to a preset quantity. That is, if the quantity of interface identifications saved in said identification sequence reaches said preset quantity, whenever a new interface identification is added, an interface identification of which the acquiring time (in another word, the save time) is the earliest will be discarded.
- the quantity of theinterfaceidentifications saved in saididentification sequence is maintained in said preset quantity, such as 25, and the most recently acquired 25 of interface identifications will always be saved.
- the interface identifications saved in the identification sequence may be the same, that is, said identification sequence may contain multiple of the same interface identification.
- S104 identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
- the more often the same interface identifications is saved in said identification sequence means that the more often the audio frequency with the maximum value is input into the audio input interface corresponding to said interface identification, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user, and the correct chance will be high to identify said audio input interface corresponding to said interface identification as a valid audio input interface.
- the accuracy of identifying will be further increased.
- said audio input interface may be automatically matched to an audio software in the background for processing.
- the input signal of said audio input interface may be subject to processing such as filtering and so on, before it is output into said audio software in the background for processing.
- a user interface may be further displayed, on which the audio input interfaceidentified currently or the microphone or other audio input devices connected to said audio input interface will be shown.
- the start and the end of the method for detecting an audio input interface of the present disclosure may be triggered through many different modes, for example, it may be set to start the detection when it is detected that the microphone is inserted, or it may be set to start the detection when a start instruction is received; also for example, it may be set to end the detection when the right input interface is detected, or it may be set to end the detection when the microphone is removed, or it may be set to end the detection when an end instruction is received, and so on.
- Fig. 2 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the second embodiment of the present disclosure
- the step S102 is specified as following sub-steps:
- the method for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 1 in that, prior to said acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values (S1026), the method further comprises step S1024.
- the interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired only when it is judged that said maximum energy value among the energy values of every one of said input signals is greater than or equal to a preset energy value, otherwise, the input signals of every one of said audio input interfaces are judged as all invalid.
- the audio input with the maximum energy value is created by a noise, said audio input will be judged as an invalid signal as long as the intensity of noise is less than said preset energy value, and theresult of the identificationwill not be affected, the influence of the noise on the result of the identification is effectively reduced.
- Fig. 3 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the third embodiment of the present disclosure
- the method for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 1 in that, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step S105:
- step S102 judging the input signals of every one of audio input interfaces as valid, and going to step S102 to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
- Said VAD detection is so called voice activity detection, which can effectively detect the activity of the input signals, identify the input signal which may be the audio input, and increase the speed of identifying the activities audio input interface. If all results of the VAD detection are 0, that means every one of said audio input interfaces are currently in a muted state; if at least one of results of said audio input interfaces is 1, then at least one of said audio input interfaces has an audio input, and the input signals of every one of audio input interfaces collected herein can be judged as valid, and then the next step (acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values) is executed.
- Fig. 4 is a schematic flow diagram illustrating the method for detecting an audio input interface according to the fourth embodiment of the present disclosure
- the method for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 1 in that, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step S106:
- step S102 judging every one of said input signals as valid, and going to step S102 to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
- the so called signal-noise ratio is a ratio of a normal sound signal to a noise signal (power) with no-sound signal, and it is often represented in the unit of dB.
- the signal-noise ratios of every one of audio input signals are detected, and the input signals of the audio input interface are judged as valid only when at least one of the signal-noise ratios of said audio input interfaces is no less than the preset value of signal-noise ratio, otherwise, the input signals of the audio input interface are judged as invalid.
- the influence of the noise on identifying an active audio input interface is reduced, and the accuracy of identifying is improved.
- any two or any three of steps S1021, S105 and S106 can be selected and executed in combination simultaneously, thereby further improving the accuracy and efficiency of identifying.
- Said post-processing comprises: executing a self-adaptive microphone volume adjustment to said valid audio input interface, that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state; executing a signal-noise ratio detection, starting a noise suppression according to the result of the detection, and so on.
- a self-adaptive microphone volume adjustment to said valid audio input interface that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state
- AGC Automatic Gain Control
- the user does not need to participate in setting the device, but the device will automatically select the microphone, to which the user is inputting voice signals, and the problem of no sound will never happen, and the corresponding result of the configuration will be displayed in a user interface.
- the user simply needs to speak to the microphone he wants, then the automatically switching of the audio input interfaces will be achieved, without the need of manual setting. If the microphone is broken and no sound is collected, the device can also automatically switch off the corresponding audio input interface, without the need of manual setting.
- the user who doesn't know how to set the microphone is relieved from the difficulty, which is a convenience for the user.
- said system for detecting an audio input interface of present disclosure comprises:
- an input detecting module 11 configured to acquire input signals of every one of audio input interfaces
- an energy detecting module 12 configured to acquire signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values;
- an interface identification acquiring module 13 configured to add said interface identification acquired herein into an identification sequence in an order by acquiring time;
- an identifying module 14 configured to the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
- input signals of every one of audio input interfaces are acquired at a preset time interval, energy values of said input signals of every one of said audio input interfaces are detected, and an interface identification of the audio input interface of which the energy value is the maximum among the energy values is acquired.
- Each of the interface identifications acquired is added to a preset identification sequence in an order by acquiring time. If the same interface identification is saved more often in said identification sequence, it means that, there are more such situations that the energy value of the input signal of the audio input interface corresponding to said interface identification is the maximum, and the audio input interface corresponding to said interface identification will be identified as a valid audio input interface.
- the system of the present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, and the user does not need to switch manually, which is very convenient, moreover, the influence of the noise on the result of identifying is reduced, thereby the result of identifying is more accurate.
- the input signals of every one of audio input interfaces can be acquired through monitoring for and collecting data from each of said audio input interfaces.
- the input signals of every one of said audio input interfaces comprises input signals from a microphone hardware-connected to any one of said audio input interfaces, noise signals, and so on.
- all of the audio input interfaces can be enumerated by means of calling a function of Dsound API (Direct Sound Capture Enumerate( )).
- the input data of every one of the audio input interfaces is acquired by means of collecting the input signals of every one of said audio input interfaces.
- the parameter of each of the audio input interfaces is preset to unify the audio collecting format for each of the audio input interfaces, such as using an audio collecting format with mono-channel, and 44.1 KHz of sampling rate.
- said input detecting module 11 comprises following sub-modules:
- a collecting unit configured to simultaneously collecting input signals of every one of said audio input interfaces ;
- an encapsulating unit configured to encapsulate said input signals of every one of said audio input interfaces being collected at the same time into a frame of detection data
- an extracting unit configured to de-interleave each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces contained in said frame of detection data.
- said input detecting module 11 further comprises a saving unit, configured to save each of frames of said detection data; said extracting unit is further configured to extract one of said frames of detection data every preset frames, and then de-interleave said frame extracted.
- each of the input signals is collected in an unit of 20 milliseconds, and then signals of every one of the input signals are put in a buffer with a M*20 ( M represents the number of the devices enumerated ) milliseconds length, the corresponding data is acquired from said buffer and encapsulated.
- the input signals (such as N channels of input signals, N is a nature number) of every one of audio input interfaces are encapsulated in respective frames of detection data.
- the purpose of sampling said input signals at a preset time interval can be realized.
- the input signals of every one of audio input interfaces can be re-acquired, which is very convenient.
- each of the input signals is pre-processed so as to ensure the accuracy of the result of the detection.
- the pre-processing comprises high-pass filtering, filtering certain frequency interferences, noise suppression, and so on, so as to reduce the influence of the noise on the detection of the input signals.
- each energy value represents the intensity of a respective input signal.
- the energy value of an input signal is the maximum among the energy values, means that the intensity of the input signal is the strongest, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user.
- the interface identification of said audio input interface is acquired to identify the audio input interface of the input signal whose intensity is the strongest in the present detection.
- said interface identification acquired herein is added into an identification sequence in an order by acquiring time.
- Said identification sequence may be created in a buffer or in other types of memories, so that it is easy to be accessed.
- said identification sequence follows the rule of first-in-and-first-out in an order by acquiring time, and the quantity of the interface identifications saved each time is less than or equal to a preset quantity. That is, if the quantity of interface identifications saved in said identification sequence reaches said preset quantity, whenever a new interface identification is added, an interface identification of which the acquiring time (in another word, the save time) is the earliest will be discarded.
- the quantity of the interface identifications saved in said identification sequence is maintained in said preset quantity, such as 25, and the most recently acquired 25 of interface identifications will always be saved.
- the interface identifications saved in the identification sequence may be the same, that is, said identification sequence may contain multiple of the same interface identification.
- identifying module 14 it is configured to identify the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
- the more often the same interface identifications is saved in said identification sequence means that the more often the audio frequency with the maximum value is input into the audio input interface corresponding to said interface identification, and the audio input interface corresponding to said input signal is most likely connected to a microphone being used by the user, and the correct chance will be high to identify said audio input interface corresponding to said interface identification as a valid audio input interface. And with the increasing of the quantity of the interface identifications saved in said identification sequence, the accuracy of identifying will be further increased.
- said audio input interface may be automatically matched to an audio software in the background for processing.
- the input signal of said audio input interface may be subject to processing such as filtering and so on, before it is output into said audio software in the background for processing.
- a user interface may be further displayed, on which the audio input interface identified currently or the microphone or other audio input devices connected to said audio input interface will be shown.
- the start and the end of the system for detecting an audio input interface of the present disclosure may be triggered through many different modes, for example, it may be set to start the detection when it is detected that the microphone is inserted, or it may be set to start the detection when a start instruction is received; also for example, it may be set to end the detection when the right input interface is detected, or it may be set to end the detection when the microphone is removed, or it may be set to end the detection when an end instruction is received, and so on.
- the present disclosure further provides a second embodiment of the system for detecting an audio input interface, which is mainly different from the first embodiment as shown in Fig. 5 in that, prior to acquiring an interface identification of said audio input interface of which the energy value is the maximum among the energy values, said interface identification acquiring module 13 further judges whether said maximum energy value among the energy values of every one of said input signals is no less than (greater than or equal to) a preset energy value; if it is, then acquires the interface identification of the audio input interface of which the energy value is the maximum among the energy values; if it is not, then judges the input signals of every one of the audio input interfaces as all invalid, and re-acquires input signals of every one of said input interfaces;
- said interface identification acquiring module 13 acquires the interface identification of the audio input interface of which the energy value is the maximum among the energy values only when it is judged that said maximum energy value among the energy values of every one of said input signals is greater than or equal to a preset energy value, otherwise, the input signals of every one of said audio input interfaces are judged as all invalid.
- said audio input with the maximum energy value is created by a noise, said audio input will be judged as an invalid signal as long as the intensity of noise is less than said preset energy value, and the result of the identification will not be affected, the influence of the noise on the result of the identification is effectively reduced.
- Fig. 6 is a structure diagram illustrating the system for detecting an audio input interface according to the third embodiment of the present disclosure
- the system for detecting an audio input interface in this embodiment is mainly different from the first embodiment as shown in Fig. 5 in that, the system further comprises a VAD detecting module 15;
- Said VAD detecting module 15 is configured to execute VAD detection on the input signals of every one of said audio input interfaces;
- Said VAD detection is so called voice activity detection, which can effectively detect the activity of the input signals, identify the input signal which may be the audio input, and increase the speed of identifying the activities audio input interface. If all results of the VAD detection are 0, that means every one of said audio input interfaces are currently in a muted state; if at least one of results of said audio input interfaces is 1, then at least one of said audio input interfaces has an audio input, and the input signals of every one of audio input interfaces collected herein can be judged as valid, and then the next step (acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values) is executed.
- the system is mainly different from the first embodiment as shown in Fig. 5 in that, the system further comprises a signal-noise ratio detecting module 16, which is configured to acquire signal-noise ratios of the input signals of every one of said audio input interfaces; if each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then judging the input signals of every one of said audio input interfaces as invalid, and re-acquiring input signals of every one of said audio input interfaces; if at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then judging every one of said input signals as valid, said energy detecting module 12 acquires energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the
- the so called signal-noise ratio is a ratio of a normal sound signal to a noise signal (power) with no-sound signal, and it is often represented in the unit of dB.
- the signal-noise ratios of every one of audio input signals are detected, and the input signals of the audio input interface are judged as valid only when at least one of the signal-noise ratios of said audio input interfaces is no less than the preset value of signal-noise ratio, otherwise, the input signals of the audio input interface are judged as invalid.
- the influence of the noise on identifying an active audio input interface is reduced, and the accuracy of identifying is improved.
- any two or any three of said interface identification acquiring module 13, said VAD detecting module 15 and said signal-noise ratio detecting module 16 can be selected and adopted in combination, thereby further improving the accuracy and efficiency of identifying.
- post-processing is further executed to the audio input interface identified by adjusting parameters related to the device so as to make the microphone connected to said device in the best working state.
- Said post-processing comprises: executing a self-adaptive microphone volume adjustment to said valid audio input interface, that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state; executing a signal-noise ratio detection, starting a noise suppression according to the result of the detection, and so on.
- a self-adaptive microphone volume adjustment to said valid audio input interface that is, software/hardware AGC (Automatic Gain Control) processing, self-adaptively adjusting the volume of the microphone, so as to make the microphone in the best volume state
- AGC Automatic Gain Control
- the user does not need to participate in setting the device, but the device will automatically select the microphone, to which the user is inputting voice signals, and the problem of no sound will never happen, and the corresponding result of the configuration will be displayed in a user interface.
- the user simply needs to speak to the microphone he wants, then the automatically switching of the audio input interfaces will be achieved, without the need of manual setting. If the microphone is broken and no sound is collected, the device can also automatically switch off the corresponding audio input interface, without the need of the user's manual setting.
- the user who doesn't know how to set the microphone is relieved from the difficulty, which is a convenience for the user.
- the person skilled in the art can understand that, all of or part of the processes and the corresponding system implementing the embodiments mentioned above, may be achieved by means of relevant hardware commanded by computer programs, the computer programs may be saved in the computer readable storage medium, and they may comprise the processes of embodiments of the respective methods and systems mentioned above when the programs are executed.
- the storage medium may be a disk or CD or read-only memory or random access memory, and so on.
Landscapes
- Engineering & Computer Science (AREA)
- General Health & Medical Sciences (AREA)
- Health & Medical Sciences (AREA)
- Physics & Mathematics (AREA)
- Theoretical Computer Science (AREA)
- Acoustics & Sound (AREA)
- Signal Processing (AREA)
- Otolaryngology (AREA)
- Multimedia (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Human Computer Interaction (AREA)
- General Engineering & Computer Science (AREA)
- General Physics & Mathematics (AREA)
- Circuit For Audible Band Transducer (AREA)
- User Interface Of Digital Computer (AREA)
- Telephone Function (AREA)
Abstract
Provided is a method for detecting an audio input interface, comprising: acquiring input signals of every one of audio input interfaces; acquiring energy values of said input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values; adding said interface identification acquired herein into an identification sequence in an order by acquiring time; and identifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface. Also provided is a system for detecting an audio input interface. The present disclosure is capable of effectively identifying the audio input interface connected to the microphone that the user is using, so that the user does not need to switch manually.
Description
【FIELD OF THE INVENTION】
This application claims priority to Chinese
Application No. 201310202043.3, entitled 'Method and System for Detecting an
Audio Input Interface', filed on May 27, 2013, which is hereby incorporated by
reference in its entirety.
The present disclosure relates to the field of audio
processing , and more particularly to a method and a device for detecting an
audio input interface.
【BACKGROUND OF THE INVENTION】
Along with the popularity ofvoice software, it is
becoming more accepted by most computer users, and has been an indispensable
part in people's daily life gradually. The existing computer device often
provides an option for choosing an audio input interface, which needs the user
to manually switch to choose among different audio input interfaces, however,
the switching method herein requires the user to manually try to choose each
audio input interface one by one until a voice signal is heard, which is very
inconvenient. Moreover, the user often makes misconnection due to not knowing
the correct audio input interface, as a result, the correct voice input can not
be acquired.
【SUMMARY OF THE INVENTION】
In view of the defects existing in conventional
method and device mentioned above that, the microphone, to which a user is
inputting voice signals, can not be identified automatically, but the user
needs to manually switch to choose among audio input interfaces one by one,
which is very inconvenient, an object of the present disclosure is to provide a
method for detecting an audio input interface, by means of detecting the input
data of every one of audio input interfaces, the audio input interface
connected to the microphone, into which a voice signal is being input, can be
effectively identified, which does not need the user to switch manually and is
very convenient.
In one aspect, the present disclosure is realized by
the following technical scheme:
A method for detecting an audio input interface
comprises:
acquiring input signals of every one of audio input
interfaces;
acquiring energy values of said input signals of
every one of said audio input interfaces, and acquiring an interface
identification of the audio input interface of which the energy value is the
maximum among the energy values;
adding said interface identification acquired herein
into an identification sequence in an order by acquiring time; and
identifying the audio input interface, of which the
interface identification is saved most often in said identification sequence,
as a valid audio input interface.
In another aspect, the present disclosure is to
provide a system for detecting an audio input interface, comprising:
an input detecting module, configured to acquire
input signals of every one of audio input interfaces;
an energy detecting module, configured to acquire
energy values of the input signals of every one of said audio input interfaces,
and acquire an interface identification of the audio input interface of which
the energy value is the maximum among the energy values;
an interface identification acquiring module,
configured to add said interface identification acquired herein into an
identification sequence in an order by acquiring time; and
an identifying module, configured to identify the
audio input interface, of which the interface identification is saved most
often in said identification sequence, as a valid audio input interface.
According to the method and system for detecting an
audio input interface of the present disclosure, input signals of every one of
audio input interfaces are acquired at a preset time interval, energy values of
said input signals of every one of said audio input interfaces are detected,
and an interface identification of the audio input interface of which the
energy value is the maximum among the energy values is acquired. Each of the
interface identifications acquired is added to a preset identification sequence
in an order by acquiring time. If the same interface identification is saved
more often in said identification sequence, it means that, there are more such
situations that the energy value of the input signal of the audio input
interface corresponding to said i nterface identification is the maximum, and
the audio input interface corresponding to said interface identification will
be identified as a valid audio input interface. The present disclosure is
capable of effectively identifying the audio input interface connected to the
microphone that the user is using, and the user does not need to switch
manually, which is very convenient, moreover, the influence of the noise on the
result of identifying is reduced, thereby the result of identifying is more
accurate .
【BRIEF DESCRIPTION OF THE DRAWINGS】
FIG. 1 is a schematic flow diagram illustrating the
method for detecting an audio input interface according to the first embodiment
of the present disclosure ;
FIG. 2 is a schematic flow diagram illustrating the
method for detecting an audio input interface according tothe second embodiment
of the present disclosure;
FIG. 3 is a schematic flow diagram illustrating the
method for detecting an audio input interface according to the third embodiment
of the present disclosure ;
FIG. 4 is a schematic flow diagram illustrating the
method for detecting an audio input interface according to the fourth
embodiment of the present disclosure;
FIG. 5 is a structure diagram illustrating the
system for detecting an audio input interface according to the first embodiment
of the present disclosure;
FIG. 6 is a structure diagram illustrating the
system for detecting an audio input interface according to the third embodiment
of the present disclosure;
FIG. 7 is a structure diagram illustrating the
system for detecting an audio input interface according to the fourth
embodiment of the present disclosure.
【DETAILED DESCRIPTION OF THE EMBODIMENTS】
In order to make the purpose, technical solutions
and advantages of the present disclosure to be understood more clearly, the
present disclosure will be described in further details with the accompanying
drawings and the following embodiments. It should be understood that the
specific embodiments described herein are merely examples to illustrate the
disclosure, not to limit the present disclosure.
As shown in Fig. 1, which is a schematic flow
diagram illustrating the method for detecting an audio input interface
according to the first embodiment of the present disclosure, the method for
detecting an audio input interface of said embodiment comprises following
steps:
S101: acquiring input signals of every one of audio
input interfaces ;
S102: acquiring energy values of the input signals
of every one of said audio input interfaces, and acquiring an interface
identification of the audio input interface of which the energy value is the
maximum among the energy values;
S103: adding said interface identification acquired
herein into a preset identification sequence in an order by acquiring time
;
S104: identifying the audio input interface, of
which the interface identification is saved most often in said identification
sequence, as a valid audio input interface.
According to the method for detecting an audio input
interface of the present disclosure, input signals of every one of audio input
interfaces are acquired at a preset time interval, energy values of said input
signals of every one of said audio input interfaces are detected, and an
interface identification of the audio input interface of which the energy value
is the maximum among the energy values is acquired. Each of the interface
identifications acquired is added to a preset identification sequence in an
order by acquiring time. If the same interface identification is saved more
often in said identification sequence, it means that, there are more such
situations that the energy value of the input signal of the audio input
interface corresponding to said interface identification is the maximum, and
the audio input interface corresponding to said interface identification will
be identified as a valid audio input interface. The method of the present
disclosure is capable of effectively identifying the audio input interface
connected to the microphone that the user is using, and the user does not need
to switch manually, which is very convenient, moreover, the influence of the
noise on the result of identifying is reduced, thereby the result of
identifying is more accurate.
Wherein, with respect to step S101 , the input
signals of every one of audio input interfaces can be acquired through
monitoring for and collecting data from each of saidaudio input interfaces. The
input signals of every one of said audio input interfaces comprise input
signals from a microphone hardware-connected to any one of said audio input
interfaces, noise signals, and so on. In one embodiment, all of the audio input
interfaces can be enumerated by means of calling a function of Dsound API
(Direct Sound Capture Enumerate( )).
After all of the audio input interfaces of the
device having been detected, the input data of every one of theaudio input
interfaces is acquired by means of collecting the input signals of every one of
said audio input interfaces. Preferably, the parameter of each of the audio
input interfaces is preset to unify the audio collecting format for each of the
audio input interfaces, such as using an audio collecting format with
mono-channel, and 44.1 KHz of sampling rate. By unifying the audio collecting
format of each of the audio input interfaces, a large amount of computational
load can be reduced in the post processing of the input signals, and the speed
of identifying the microphone can be increased.
In one embodiment, input signals of every one of
audio input interfaces are acquired simultaneously at a preset time interval,
which comprises following sub-steps:
S1011: simultaneously collecting input signals of
every one of said audio input interfaces, and encapsulating said input signals
of every one of said audio input interfaces collected at the same time into a
frame of detection data.
S1012: de-interleaving each of frames of said
detection data so as to acquire the input signals of every one of said audio
input interfaces. Moreover, each of frames of said detection data can be saved;
one of said frames of detection data is extracted every preset frames and
further de-interleaved so as to acquire the input signals of every one of said
audio input interfaces contained in said frame of detection data .
In one embodiment, whenencapsulatingthe input
signals of every one of said audio input interfaces, each of the input signals
is collected in an unit of 20 milliseconds, and then signals of every one of
the input signals are put in a buffer with a M*20(M represents the number of
the devices enumerated)milliseconds length, the corresponding data is acquired
from said buffer and encapsulated.
In this way, the input signals (such as N channels
of input signals, N is a nature number) of every one of audio input interfaces
are encapsulated in respective frames of detection data. By simply extracting
one frame of said detection data every certain frames, the purpose of sampling
said input signals at a preset time interval can be realized. By simply
de-interleaving the detection data extracted, the input signals of every one of
audio input interfaces can be re-acquired, which is very convenient.
Furthermore, after the input signals of every one of
said audio input interfaces have been acquired, each of the input signals is
pre-processed so as to ensure the accuracy of the result of the detection. The
pre-processing comprises high-pass filtering, filtering certain frequency
interferences, noise suppression, and so on, so as to reduce the influence of
the noise on the detection of the input signals.
With respect to S102, the energy values of the input
signals of every one of said audio input interfaces acquired herein are detect
ed and acquired, each energy value represents the intensity of a respective
input signal. The energy value of an input signal is the maximum among the
energy values, means that the intensity of the input signal is the strongest,
and the audio input interface corresponding to said input signal is most likely
connected to a microphone being used by the user. The interface identification
of said audio input interface is acquired to identify the audio input interface
of the input signal whose intensity is the strongest in the present
detection.
With respect to S103, said interface identification
acquired herein is added into an identification sequence in an order by
acquiring time. Said identification sequence may be created in a buffer or in
other types of memories, so that it is easy to be accessed.
Preferably, said identification sequence follows the
rule of first-in-and-first-out in an order by acquiring time, and the quantity
of the interface identifications saved each time is less than or equal to a
preset quantity. That is, if the quantity of interface identifications saved in
said identification sequence reaches said preset quantity, whenever a new
interface identification is added, an interface identification of which the
acquiring time (in another word, the save time) is the earliest will be
discarded. The quantity of theinterfaceidentifications saved in
saididentification sequence is maintained in said preset quantity, such as 25,
and the most recently acquired 25 of interface identifications will always be
saved. The interface identifications saved in the identification sequence may
be the same, that is, said identification sequence may contain multiple of the
same interface identification.
S104: identifying the audio input interface, of
which the interface identification is saved most often in said identification
sequence, as a valid audio input interface.
The more often the same interface identifications is
saved in said identification sequence, means that the more often the audio
frequency with the maximum value is input into the audio input interface
corresponding to said interface identification, and the audio input interface
corresponding to said input signal is most likely connected to a microphone
being used by the user, and the correct chance will be high to identify said
audio input interface corresponding to said interface identification as a valid
audio input interface. And with the increasing of the quantity of the interface
identifications saved in said identification sequence, the accuracy of
identifying will be further increased. Moreover, after the valid audio input
interface has been identified, said audio input interface may be automatically
matched to an audio software in the background for processing. Alternatively,
the input signal of said audio input interface may be subject to processing
such as filtering and so on, before it is output into said audio software in
the background for processing. After said audio input interface has been
identified, a user interface may be further displayed, on which the audio input
interfaceidentified currently or the microphone or other audio input devices
connected to said audio input interface will be shown. The start and the end of
the method for detecting an audio input interface of the present disclosure may
be triggered through many different modes, for example, it may be set to start
the detection when it is detected that the microphone is inserted, or it may be
set to start the detection when a start instruction is received; also for
example, it may be set to end the detection when the right input interface is
detected, or it may be set to end the detection when the microphone is removed,
or it may be set to end the detection when an end instruction is received, and
so on.
As shown in Fig. 2, which is a schematic flow
diagram illustrating the method for detecting an audio input interface
according to the second embodiment of the present disclosure, the step S102 is
specified as following sub-steps:
S1022 : acquiring energy values of the input signals
of every one of said audio input interfaces ;
S1024: judging whether said maximum energy value
among the energy values of every one of said input signals is no less than
(greater than or equal to) a preset energy value;
if it is, then going to S1026; if it is not, then
judging the input signals of every one of the audio input interfaces as all
invalid, and re-acquiring input signals of every one of said input
interfaces;
S1026: acquiring the interface identification of the
audio input interface of which the energy value is the maximum among the energy
values.
The method for detecting an audio input interface in
this embodiment is mainly different from the first embodiment as shown in Fig.
1 in that, prior to said acquiring an interface identification of the audio
input interface of which the energy value is the maximum among the energy
values (S1026), the method further comprises step S1024.
According to the method for detecting an audio input
interface in this embodiment, the interface identification of the audio input
interface of which the energy value is the maximum among the energy values is
acquired only when it is judged that said maximum energy value among the energy
values of every one of said input signals is greater than or equal to a preset
energy value, otherwise, the input signals of every one of said audio input
interfaces are judged as all invalid. As a result, if the audio input with the
maximum energy value is created by a noise, said audio input will be judged as
an invalid signal as long as the intensity of noise is less than said preset
energy value, and theresult of the identificationwill not be affected, the
influence of the noise on the result of the identification is effectively
reduced.
As shown in Fig. 3, which is a schematic flow
diagram illustrating the method for detecting an audio input interface
according to the third embodiment of the present disclosure, the method for
detecting an audio input interface in this embodiment is mainly different from
the first embodiment as shown in Fig. 1 in that, after said acquiring input
signals of every one of audio input interfaces, the method further comprises
the following step S105:
S105: executing VAD detection on the input signals
of every one of said audio input interfaces;
if all results of said VAD detection on every one of
said audio input interfaces are zero (0) , then judging said input signals of
every one of said audio input interfaces as all invalid, and re-acquiring input
signals of every one of said input interfaces ;
if at least one of the results of said VAD detection
of every one of said audio input interfaces is one (1), then judging the input
signals of every one of audio input interfaces as valid, and going to step S102
to acquire energy values of the input signals of every one of said audio input
interfaces, and acquire an interface identification of the audio input
interface of which the energy value is the maximum among the energy values.
Said VAD detection is so called voice activity
detection, which can effectively detect the activity of the input signals,
identify the input signal which may be the audio input, and increase the speed
of identifying the activities audio input interface. If all results of the VAD
detection are 0, that means every one of said audio input interfaces are
currently in a muted state; if at least one of results of said audio input
interfaces is 1, then at least one of said audio input interfaces has an audio
input, and the input signals of every one of audio input interfaces collected
herein can be judged as valid, and then the next step (acquiring energy values
of the input signals of every one of said audio input interfaces, and acquiring
an interface identification of the audio input interface of which the energy
value is the maximum among the energy values) is executed.
As shown in Fig. 4, which is a schematic flow
diagram illustrating the method for detecting an audio input interface
according to the fourth embodiment of the present disclosure, the method for
detecting an audio input interface in this embodiment is mainly different from
the first embodiment as shown in Fig. 1 in that, after said acquiring input
signals of every one of audio input interfaces, the method further comprises
the following step S106:
acquiring signal-noise ratios of the input signals
of every one of said audio input interfaces ;
if each of said signal-noise ratios of said input
signals is less than a preset value of signal-noise ratio, then judging the
input signals of every one of said audio input interfaces as invalid, and
re-acquiring input signals of every one of said audio input interfaces;
if at least one of said signal-noise ratios of each
of said input signals is no less than said preset value of signal-noise ratio,
then judging every one of said input signals as valid, and going to step S102
to acquire energy values of the input signals of every one of said audio input
interfaces, and acquire an interface identification of the audio input
interface of which the energy value is the maximum among the energy values.
The so called signal-noise ratio is a ratio of a
normal sound signal to a noise signal (power) with no-sound signal, and it is
often represented in the unit of dB. In this embodiment, the signal-noise
ratios of every one of audio input signals are detected, and the input signals
of the audio input interface are judged as valid only when at least one of the
signal-noise ratios of said audio input interfaces is no less than the preset
value of signal-noise ratio, otherwise, the input signals of the audio input
interface are judged as invalid. As a result, the influence of the noise on
identifying an active audio input interface is reduced, and the accuracy of
identifying is improved.
In a preferred embodiment, any two or any three of
steps S1021, S105 and S106 can be selected and executed in combination
simultaneously, thereby further improving the accuracy and efficiency of
identifying.
In another preferred embodiment of the method, after
said identifying the audio input interface, of which the interface
identification is saved most often in said identification sequence, as a valid
audio input interface, further executing a step of post-processing to the audio
input interface identified by adjusting parameters related to the device so as
to make the microphone connected to said device in the best working state.
Said post-processing comprises: executing a
self-adaptive microphone volume adjustment to said valid audio input interface,
that is, software/hardware AGC (Automatic Gain Control) processing,
self-adaptively adjusting the volume of the microphone, so as to make the
microphone in the best volume state; executing a signal-noise ratio detection,
starting a noise suppression according to the result of the detection, and so
on.
According to the method of the present disclosure,
the user does not need to participate in setting the device, but the device
will automatically select the microphone, to which the user is inputting voice
signals, and the problem of no sound will never happen, and the corresponding
result of the configuration will be displayed in a user interface. When there
are multiple microphones with different acoustic characteristics connected to
one device, the user simply needs to speak to the microphone he wants, then the
automatically switching of the audio input interfaces will be achieved, without
the need of manual setting. If the microphone is broken and no sound is
collected, the device can also automatically switch off the corresponding audio
input interface, without the need of manual setting. With the method of the
present disclosure, the user who doesn't know how to set the microphone is
relieved from the difficulty, which is a convenience for the user.
As shown in Fig. 5, which is a structure diagram
illustrating the system for detecting an audio input interface according to the
first embodiment of the present disclosure, said system for detecting an audio
input interface of present disclosure comprises:
an input detecting module 11, configured to acquire
input signals of every one of audio input interfaces;
an energy detecting module 12, configured to acquire
signals of every one of said audio input interfaces, and acquire an interface
identification of the audio input interface of which the energy value is the
maximum among the energy values;
an interface identification acquiring module 13,
configured to add said interface identification acquired herein into an
identification sequence in an order by acquiring time;
an identifying module 14, configured to the audio
input interface, of which the interface identification is saved most often in
said identification sequence, as a valid audio input interface.
According to the system for detecting an audio input
interface of the present disclosure, input signals of every one of audio input
interfaces are acquired at a preset time interval, energy values of said input
signals of every one of said audio input interfaces are detected, and an
interface identification of the audio input interface of which the energy value
is the maximum among the energy values is acquired. Each of the interface
identifications acquired is added to a preset identification sequence in an
order by acquiring time. If the same interface identification is saved more
often in said identification sequence, it means that, there are more such
situations that the energy value of the input signal of the audio input
interface corresponding to said interface identification is the maximum, and
the audio input interface corresponding to said interface identification will
be identified as a valid audio input interface. The system of the present
disclosure is capable of effectively identifying the audio input interface
connected to the microphone that the user is using, and the user does not need
to switch manually, which is very convenient, moreover, the influence of the
noise on the result of identifying is reduced, thereby the result of
identifying is more accurate.
Wherein, with respect to said input detecting module
11, the input signals of every one of audio input interfaces can be acquired
through monitoring for and collecting data from each of said audio input
interfaces. The input signals of every one of said audio input interfaces
comprises input signals from a microphone hardware-connected to any one of said
audio input interfaces, noise signals, and so on. In one embodiment, all of the
audio input interfaces can be enumerated by means of calling a function of
Dsound API (Direct Sound Capture Enumerate( )).
After all of the audio input interfaces of the
device having been detected, the input data of every one of the audio input
interfaces is acquired by means of collecting the input signals of every one of
said audio input interfaces. Preferably, the parameter of each of the audio
input interfaces is preset to unify the audio collecting format for each of the
audio input interfaces, such as using an audio collecting format with
mono-channel, and 44.1 KHz of sampling rate. By unifying the audio collecting
format of each of the audio input interfaces, a large amount of computational
load can be reduced in the post processing of the input signals, and the speed
of identifying the microphone can be increased.
In one embodiment, said input detecting module 11
comprises following sub-modules:
a collecting unit, configured to simultaneously
collecting input signals of every one of said audio input interfaces ;
an encapsulating unit, configured to encapsulate
said input signals of every one of said audio input interfaces being collected
at the same time into a frame of detection data;
an extracting unit, configured to de-interleave each
of frames of said detection data so as to acquire the input signals of every
one of said audio input interfaces contained in said frame of detection
data.
Moreover, said input detecting module 11 further
comprises a saving unit, configured to save each of frames of said detection
data; said extracting unit is further configured to extract one of said frames
of detection data every preset frames, and then de-interleave said frame
extracted.
For example, when the encapsulating unit
encapsulates the input signals of every one of said audio input interfaces,
each of the input signals is collected in an unit of 20 milliseconds, and then
signals of every one of the input signals are put in a buffer with a M*20 ( M
represents the number of the devices enumerated ) milliseconds length, the
corresponding data is acquired from said buffer and encapsulated.
In this way, the input signals (such as N channels
of input signals, N is a nature number) of every one of audio input interfaces
are encapsulated in respective frames of detection data. By simply extracting
one frame of said detection data every certain frames, the purpose of sampling
said input signals at a preset time interval can be realized. By simply
de-interleaving the detection data extracted, the input signals of every one of
audio input interfaces can be re-acquired, which is very convenient.
Furthermore, after the input detecting module 11 has
acquired the input signals of every one of said audio input interfaces, each of
the input signals is pre-processed so as to ensure the accuracy of the result
of the detection. The pre-processing comprises high-pass filtering, filtering
certain frequency interferences, noise suppression, and so on, so as to reduce
the influence of the noise on the detection of the input signals.
With respect to said energy detecting module 12, the
energy values of the input signals of every one of said audio input interfaces
acquired herein are detected and acquired, each energy value represents the
intensity of a respective input signal. The energy value of an input signal is
the maximum among the energy values, means that the intensity of the input
signal is the strongest, and the audio input interface corresponding to said
input signal is most likely connected to a microphone being used by the user.
The interface identification of said audio input interface is acquired to
identify the audio input interface of the input signal whose intensity is the
strongest in the present detection.
With respect to said interface identification
acquiring module 13, said interface identification acquired herein is added
into an identification sequence in an order by acquiring time. Said
identification sequence may be created in a buffer or in other types of
memories, so that it is easy to be accessed.
Preferably, said identification sequence follows the
rule of first-in-and-first-out in an order by acquiring time, and the quantity
of the interface identifications saved each time is less than or equal to a
preset quantity. That is, if the quantity of interface identifications saved in
said identification sequence reaches said preset quantity, whenever a new
interface identification is added, an interface identification of which the
acquiring time (in another word, the save time) is the earliest will be
discarded. The quantity of the interface identifications saved in said
identification sequence is maintained in said preset quantity, such as 25, and
the most recently acquired 25 of interface identifications will always be
saved. The interface identifications saved in the identification sequence may
be the same, that is, said identification sequence may contain multiple of the
same interface identification.
With respect to said identifying module 14, it is
configured to identify the audio input interface, of which the interface
identification is saved most often in said identification sequence, as a valid
audio input interface.
The more often the same interface identifications is
saved in said identification sequence, means that the more often the audio
frequency with the maximum value is input into the audio input interface
corresponding to said interface identification, and the audio input interface
corresponding to said input signal is most likely connected to a microphone
being used by the user, and the correct chance will be high to identify said
audio input interface corresponding to said interface identification as a valid
audio input interface. And with the increasing of the quantity of the interface
identifications saved in said identification sequence, the accuracy of
identifying will be further increased.
Moreover, after the valid audio input interface has
been identified, said audio input interface may be automatically matched to an
audio software in the background for processing. Alternatively, the input
signal of said audio input interface may be subject to processing such as
filtering and so on, before it is output into said audio software in the
background for processing. After said audio input interface has been
identified, a user interface may be further displayed, on which the audio input
interface identified currently or the microphone or other audio input devices
connected to said audio input interface will be shown.
The start and the end of the system for detecting an
audio input interface of the present disclosure may be triggered through many
different modes, for example, it may be set to start the detection when it is
detected that the microphone is inserted, or it may be set to start the
detection when a start instruction is received; also for example, it may be set
to end the detection when the right input interface is detected, or it may be
set to end the detection when the microphone is removed, or it may be set to
end the detection when an end instruction is received, and so on.
The present disclosure further provides a second
embodiment of the system for detecting an audio input interface, which is
mainly different from the first embodiment as shown in Fig. 5 in that, prior to
acquiring an interface identification of said audio input interface of which
the energy value is the maximum among the energy values, said interface
identification acquiring module 13 further judges whether said maximum energy
value among the energy values of every one of said input signals is no less
than (greater than or equal to) a preset energy value; if it is, then acquires
the interface identification of the audio input interface of which the energy
value is the maximum among the energy values; if it is not, then judges the
input signals of every one of the audio input interfaces as all invalid, and
re-acquires input signals of every one of said input interfaces;
According to the system for detecting an audio input
interface in this embodiment, said interface identification acquiring module 13
acquires the interface identification of the audio input interface of which the
energy value is the maximum among the energy values only when it is judged that
said maximum energy value among the energy values of every one of said input
signals is greater than or equal to a preset energy value, otherwise, the input
signals of every one of said audio input interfaces are judged as all invalid.
As a result, if the audio input with the maximum energy value is created by a
noise, said audio input will be judged as an invalid signal as long as the
intensity of noise is less than said preset energy value, and the result of the
identification will not be affected, the influence of the noise on the result
of the identification is effectively reduced.
As shown in Fig. 6, which is a structure diagram
illustrating the system for detecting an audio input interface according to the
third embodiment of the present disclosure, the system for detecting an audio
input interface in this embodiment is mainly different from the first
embodiment as shown in Fig. 5 in that, the system further comprises a VAD
detecting module 15;
Said VAD detecting module 15 is configured to
execute VAD detection on the input signals of every one of said audio input
interfaces;
if all results of said VAD detection on every one of
said audio input interfaces are zero (0), then judging said input signals of
every one of said audio input interfaces as all invalid, and re-acquiring input
signals of every one of said input interfaces; if at least one of the results
of said VAD detection of every one of said audio input interfaces is one (1),
then judging the input signals of every one of audio input interfaces as valid,
and said energy detecting module 12 acquires energy values of the input signals
of every one of said audio input interfaces, and acquire an interface
identification of the audio input interface of which the energy value is the
maximum among the energy values.
Said VAD detection is so called voice activity
detection, which can effectively detect the activity of the input signals,
identify the input signal which may be the audio input, and increase the speed
of identifying the activities audio input interface. If all results of the VAD
detection are 0, that means every one of said audio input interfaces are
currently in a muted state; if at least one of results of said audio input
interfaces is 1, then at least one of said audio input interfaces has an audio
input, and the input signals of every one of audio input interfaces collected
herein can be judged as valid, and then the next step (acquiring energy values
of the input signals of every one of said audio input interfaces, and acquiring
an interface identification of the audio input interface of which the energy
value is the maximum among the energy values) is executed.
As shown in F ig . 7, which is a structure diagram
illustrating the system for detecting an audio input interface according to the
fourth embodiment of the present disclosure, the system is mainly different
from the first embodiment as shown in Fig. 5 in that, the system further
comprises a signal-noise ratio detecting module 16, which is configured to
acquire signal-noise ratios of the input signals of every one of said audio
input interfaces; if each of said signal-noise ratios of said input signals is
less than a preset value of signal-noise ratio, then judging the input signals
of every one of said audio input interfaces as invalid, and re-acquiring input
signals of every one of said audio input interfaces; if at least one of said
signal-noise ratios of each of said input signals is no less than said preset
value of signal-noise ratio, then judging every one of said input signals as
valid, said energy detecting module 12 acquires energy values of the input
signals of every one of said audio input interfaces, and acquire an interface
identification of the audio input interface of which the energy value is the
maximum among the energy values.
The so called signal-noise ratio is a ratio of a
normal sound signal to a noise signal (power) with no-sound signal, and it is
often represented in the unit of dB. In this embodiment, the signal-noise
ratios of every one of audio input signals are detected, and the input signals
of the audio input interface are judged as valid only when at least one of the
signal-noise ratios of said audio input interfaces is no less than the preset
value of signal-noise ratio, otherwise, the input signals of the audio input
interface are judged as invalid. As a result, the influence of the noise on
identifying an active audio input interface is reduced, and the accuracy of
identifying is improved.
In a preferred embodiment, any two or any three of
said interface identification acquiring module 13, said VAD detecting module 15
and said signal-noise ratio detecting module 16 can be selected and adopted in
combination, thereby further improving the accuracy and efficiency of
identifying.
In another preferred embodiment of the system,
after said identifying module 14 identifies the audio input interface, of which
the interface identification is saved most often in said identification
sequence, as a valid audio input interface, post-processing is further executed
to the audio input interface identified by adjusting parameters related to the
device so as to make the microphone connected to said device in the best
working state.
Said post-processing comprises: executing a
self-adaptive microphone volume adjustment to said valid audio input interface,
that is, software/hardware AGC (Automatic Gain Control) processing,
self-adaptively adjusting the volume of the microphone, so as to make the
microphone in the best volume state; executing a signal-noise ratio detection,
starting a noise suppression according to the result of the detection, and so
on.
According to the system of the present disclosure,
the user does not need to participate in setting the device, but the device
will automatically select the microphone, to which the user is inputting voice
signals, and the problem of no sound will never happen, and the corresponding
result of the configuration will be displayed in a user interface. When there
are multiple microphones with different acoustic characteristics connected to
one device, the user simply needs to speak to the microphone he wants, then the
automatically switching of the audio input interfaces will be achieved, without
the need of manual setting. If the microphone is broken and no sound is
collected, the device can also automatically switch off the corresponding audio
input interface, without the need of the user's manual setting. With the system
of the present disclosure, the user who doesn't know how to set the microphone
is relieved from the difficulty, which is a convenience for the user.
The person skilled in the art can understand that,
all of or part of the processes and the corresponding system implementing the
embodiments mentioned above, may be achieved by means of relevant hardware
commanded by computer programs, the computer programs may be saved in the
computer readable storage medium, and they may comprise the processes of
embodiments of the respective methods and systems mentioned above when the
programs are executed. Wherein, the storage medium may be a disk or CD or
read-only memory or random access memory, and so on.
The foregoing examples are preferred embodiments of
the present disclosure only and not intended to limit the present disclosure.
It should be understood that, to the person skilled in the art, various
modifications and improvements can be made without departing from the spirit
and principle of the present disclosure, which should all be included within
the scope of the present disclosure. Therefore, the protection scope of the
present disclosure shall be defined by the appended claims.
Claims (20)
- A methodfor detecting an audio input interface , comprising :acquiring input signals of every one of audio input interfaces;acquiring energy values of said input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values;adding said interface identification acquired herein into an identification sequence in an order by acquiring time ; andidentifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
- The method for detecting an audio input interface according to claim 1, further comprising:when the quantity of interface identifications saved in said identification sequence reaches a preset quantity, whenever a new interface identification is added, one saved interface identification of which the acquiring time is the earliest will be discarded.
- The method for detecting an audio input interface according to claim 1 or 2, wherein, prior to said acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values, the method further comprises the following sub-step:judging whether said maximum energy value among the energy values of every one of said input signals is no less than a preset energy value;when said maximum energy value among the energy values of every one of said input signals is no less than the preset energy value, then acquiring the interface identification of the audio input interface of which the energy value is the maximum among the energy values; when said maximum energy value among the energy values of every one of said input signals is less than the preset energy value , then judging the input signals of every one of the audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces.
- The method for detecting an audio input interface according to claim 1 or 2, wherein, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step:executing VAD detection on the input signals of every one of said audio input interfaces;when all results of said VAD detection on every one of said audio input interfaces are zero, then judging said input signals of every one of said audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces;when at least one of the results of said VAD detection of every one of said audio input interfaces is one, then judging the input signals of every one of audio input interfaces as valid, acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
- The method for detecting an audio input interface according to claim 1 or 2, wherein, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step:acquiring signal-noise ratios of the input signals of every one of said audio input interfaces;when each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then judging the input signals of every one of said audio input interfaces as invalid, and re-acquiring input signals of every one of said audio input interfaces;when at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then judging every one of said input signals as valid, acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
- The method for detecting an audio input interface according to claim 1 or 2, wherein, said acquiring input signals of every one of audio input interfaces comprises following sub-steps:simultaneously collecting input signals of every one of said audio input interfaces, and encapsulating said input signals of every one of said audio input interfaces collected at the same time into a frame of detection data;de-interleaving each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces.
- The method for detecting an audio input interface according to claim 6, wherein, said de-interleaving each of frames of said detection data comprises the following sub-steps:saving each of frames of said detection data ;extracting one of said frames of detection data every preset frames, and then de-interleaving said frame extracted herein.
- A system for detecting an audio input interface , comprising:an input detecting module, configured to acquire input signals of every one of audio input interfaces;an energy detecting module, configured to acquire energy values of the input signals of every one of said audio input interfaces, and acquire an interface identification of the audio input interface of which the energy value is the maximum among the energy values;an interface identification acquiring module, configured to add said interface identification acquired herein into an identification sequence in an order by acquiring time; andan identifying module, configured to identify the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
- The system for detecting an audio input interface according to claim 8 , wherein, the interface identification is configured to discard one saved interface identification of which the acquiring time is the earliest whenever a new interface identification is added, if the quantity of interface identifications saved in said identification sequence reaches a preset quantity.
- The system for detecting an audio input interface according to claim 8 or 9, wherein, said interface identification acquiring module is configured to judge whether said maximum energy value among the energy values of every one of said input signals is no less than a preset energy value;when said maximum energy value among the energy values of every one of said input signals is no less than a preset energy value, then the system acquires the interface identification of the audio input interface of which the energy value is the maximum among the energy values; when said maximum energy value among the energy values of every one of said input signals is less than a preset energy value, then the system judges the input signals of every one of the audio input interfaces as all invalid, and re-acquires input signals of every one of said input interfaces.
- The system for detecting an audio input interface according to claim 8 or 9, furthering comprising:a VAD detecting module, configured to execute VAD detection on the input signals of every one of said audio input interfaces;when all results of said VAD detection on every one of said audio input interfaces are zero, then the system judges said input signals of every one of said audio input interfaces as all invalid, and re-acquires input signals of every one of said input interfaces;when at least one of the results of said VAD detection of every one of said audio input interfaces is one, then the system judges the input signals of every one of audio input interfaces as valid, the energy detecting module acquires energy values of the input signals of every one of said audio input interfaces, and acquires an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
- The system for detecting an audio input interface according to claim 8 or 9, furthering comprising :a signal-noise ratio detecting module, configured to acquire signal-noise ratios of the input signals of every one of said audio input interfaces;when each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then the system judges the input signals of every one of said audio input interfaces as invalid, and re-acquires input signals of every one of said audio input interfaces;when at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then the system judges every one of said input signals as valid, the energy detecting module acquires energy values of the input signals of every one of said audio input interfaces, and acquires an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
- The system for detecting an audio input interface according to claim 8 or 9, wherein, said input detecting module comprises following sub-modules:a collecting unit, configured to simultaneously collecting input signals of every one of said audio input interfaces ;an encapsulating unit, configured to encapsulate said input signals of every one of said audio input interfaces being collected at the same time into a frame of detection data;an extracting unit, configured to de-interleave each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces contained in said frame of detection data.
- The system for detecting an audio input interface according to claim 13, wherein, said input detecting module further comprises a saving unit, which is configured to save each of frames of said detection data; and wherein said extracting unit is further configured to extract one of said frames of detection data every preset frames, and then de-interleave said frame extracted.
- One or more non-transitory computer readable storage medium, including computer executable instructions, aid executable instructions are configured to execute a method for detecting an audio input interface, wherein, said method comprising:acquiring input signals of every one of audio input interfaces;acquiring energy values of said input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values;adding said interface identification acquired herein into an identification sequence in an order by acquiring time ; andidentifying the audio input interface, of which the interface identification is saved most often in said identification sequence, as a valid audio input interface.
- The one or more non-transitory computer readable storage medium according to claim 15, wherein, said method further comprising:when the quantity of interface identifications saved in said identification sequence reaches a preset quantity, whenever a new interface identification is added, one saved interface identification of which the acquiring time is the earliest will be discarded.
- The one or more non-transitory computer readable storage medium according to claim 16, wherein, prior to said acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values, the method further comprises the following sub-step:judging whether said maximum energy value among the energy values of every one of said input signals is no less than a preset energy value;when said maximum energy value among the energy values of every one of said input signals is no less than the preset energy value, then acquiring the interface identification of the audio input interface of which the energy value is the maximum among the energy values; when said maximum energy value among the energy values of every one of said input signals is less than the preset energy value , then judging the input signals of every one of the audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces.
- The one or more non-transitory computer readable storage medium according to claim 16, wherein, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step:executing VAD detection on the input signals of every one of said audio input interfaces;when all results of said VAD detection on every one of said audio input interfaces are zero, then judging said input signals of every one of said audio input interfaces as all invalid, and re-acquiring input signals of every one of said input interfaces;when at least one of the results of said VAD detection of every one of said audio input interfaces is one, then judging the input signals of every one of audio input interfaces as valid, acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
- The one or more non-transitory computer readable storage medium according to claim 16, wherein, after said acquiring input signals of every one of audio input interfaces, the method further comprises the following step:acquiring signal-noise ratios of the input signals of every one of said audio input interfaces;when each of said signal-noise ratios of said input signals is less than a preset value of signal-noise ratio, then judging the input signals of every one of said audio input interfaces as invalid, and re-acquiring input signals of every one of said audio input interfaces;when at least one of said signal-noise ratios of each of said input signals is no less than said preset value of signal-noise ratio, then judging every one of said input signals as valid, acquiring energy values of the input signals of every one of said audio input interfaces, and acquiring an interface identification of the audio input interface of which the energy value is the maximum among the energy values.
- The one or more non-transitory computer readable storage medium according to claim 16, wherein, said acquiring input signals of every one of audio input interfaces comprises following sub-steps:simultaneously collecting input signals of every one of said audio input interfaces, and encapsulating said input signals of every one of said audio input interfaces collected at the same time into a frame of detection data;de-interleaving each of frames of said detection data so as to acquire the input signals of every one of said audio input interfaces.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US14/683,103 US9886238B2 (en) | 2013-05-27 | 2015-04-09 | Method, system and computer storage medium for detecting an audio input interface |
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| CN201310202043.3A CN103327433B (en) | 2013-05-27 | 2013-05-27 | Audio input interface detection method and system thereof |
| CN201310202043.3 | 2013-05-27 |
Related Child Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| US14/683,103 Continuation US9886238B2 (en) | 2013-05-27 | 2015-04-09 | Method, system and computer storage medium for detecting an audio input interface |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2014190824A1 true WO2014190824A1 (en) | 2014-12-04 |
Family
ID=49195917
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/CN2014/075663 Ceased WO2014190824A1 (en) | 2013-05-27 | 2014-04-18 | Method, system and computer storage medium for detecting an audio input interface |
Country Status (3)
| Country | Link |
|---|---|
| US (1) | US9886238B2 (en) |
| CN (1) | CN103327433B (en) |
| WO (1) | WO2014190824A1 (en) |
Families Citing this family (8)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN103327433B (en) * | 2013-05-27 | 2014-08-27 | 腾讯科技(深圳)有限公司 | Audio input interface detection method and system thereof |
| WO2016015186A1 (en) * | 2014-07-28 | 2016-02-04 | 华为技术有限公司 | Acoustical signal processing method and device of communication device |
| CN105788609B (en) * | 2014-12-25 | 2019-08-09 | 福建凯米网络科技有限公司 | The correlating method and device and assessment method and system of multichannel source of sound |
| CN105681974A (en) * | 2016-04-01 | 2016-06-15 | 北京小鸟听听科技有限公司 | Multiple-sound-source switching method and device and audio equipment |
| CN111354356B (en) * | 2018-12-24 | 2024-04-30 | 北京搜狗科技发展有限公司 | A method and device for processing voice data |
| CN111736795B (en) * | 2019-06-24 | 2024-11-29 | 北京京东尚科信息技术有限公司 | Audio processing method, device, equipment and storage medium |
| CN111770427B (en) * | 2020-06-24 | 2023-01-24 | 杭州海康威视数字技术股份有限公司 | Microphone array detection method, device, equipment and storage medium |
| CN113542975A (en) * | 2021-07-02 | 2021-10-22 | 南昌华勤电子科技有限公司 | Audio signal switching circuit and electronic equipment |
Citations (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN1805489A (en) * | 2005-01-10 | 2006-07-19 | 华为技术有限公司 | Method and system of implementing report of current speaker during conference |
| US20080040117A1 (en) * | 2004-05-14 | 2008-02-14 | Shuian Yu | Method And Apparatus Of Audio Switching |
| CN102056053A (en) * | 2010-12-17 | 2011-05-11 | 中兴通讯股份有限公司 | Multi-microphone audio mixing method and device |
| CN103327433A (en) * | 2013-05-27 | 2013-09-25 | 腾讯科技(深圳)有限公司 | Audio input interface detection method and system thereof |
Family Cites Families (7)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US6956828B2 (en) * | 2000-12-29 | 2005-10-18 | Nortel Networks Limited | Apparatus and method for packet-based media communications |
| CN1271593C (en) * | 2004-12-24 | 2006-08-23 | 北京中星微电子有限公司 | Voice signal detection method |
| CN1845580A (en) * | 2005-04-07 | 2006-10-11 | 深圳Tcl新技术有限公司 | Audio and video signal source recognition and automatic switching method and apparatus |
| CN101308651B (en) * | 2007-05-17 | 2011-05-04 | 展讯通信(上海)有限公司 | Detection method of audio transient signal |
| CN101299782B (en) * | 2008-05-22 | 2011-09-21 | 杭州华三通信技术有限公司 | Method and device for detecting double-audio signal |
| US8514265B2 (en) * | 2008-10-02 | 2013-08-20 | Lifesize Communications, Inc. | Systems and methods for selecting videoconferencing endpoints for display in a composite video image |
| CN102143262B (en) * | 2010-02-03 | 2014-03-26 | 深圳富泰宏精密工业有限公司 | Electronic device and method for switching audio input channel thereof |
-
2013
- 2013-05-27 CN CN201310202043.3A patent/CN103327433B/en active Active
-
2014
- 2014-04-18 WO PCT/CN2014/075663 patent/WO2014190824A1/en not_active Ceased
-
2015
- 2015-04-09 US US14/683,103 patent/US9886238B2/en active Active
Patent Citations (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20080040117A1 (en) * | 2004-05-14 | 2008-02-14 | Shuian Yu | Method And Apparatus Of Audio Switching |
| CN1805489A (en) * | 2005-01-10 | 2006-07-19 | 华为技术有限公司 | Method and system of implementing report of current speaker during conference |
| CN102056053A (en) * | 2010-12-17 | 2011-05-11 | 中兴通讯股份有限公司 | Multi-microphone audio mixing method and device |
| CN103327433A (en) * | 2013-05-27 | 2013-09-25 | 腾讯科技(深圳)有限公司 | Audio input interface detection method and system thereof |
Also Published As
| Publication number | Publication date |
|---|---|
| CN103327433B (en) | 2014-08-27 |
| US9886238B2 (en) | 2018-02-06 |
| CN103327433A (en) | 2013-09-25 |
| US20150212792A1 (en) | 2015-07-30 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| WO2014190824A1 (en) | Method, system and computer storage medium for detecting an audio input interface | |
| US8233637B2 (en) | Multi-membrane microphone for high-amplitude audio capture | |
| US7968786B2 (en) | Volume adjusting apparatus and volume adjusting method | |
| US9020157B2 (en) | Active noise cancellation system | |
| KR101118217B1 (en) | Audio data processing apparatus and method therefor | |
| US6549630B1 (en) | Signal expander with discrimination between close and distant acoustic source | |
| EP2731351A2 (en) | Adaptive system for managing a plurality of microphones and speakers | |
| WO2020060206A1 (en) | Methods for audio processing, apparatus, electronic device and computer readable storage medium | |
| EP2426950A2 (en) | Noise suppression for sending voice with binaural microphones | |
| WO2019037319A1 (en) | Electric quantity early warning method, and server, mobile terminal and storage medium | |
| JP2012105287A (en) | Intelligibility control using ambient noise detection | |
| WO2019051890A1 (en) | Terminal control method and device, and computer-readable storage medium | |
| US9672843B2 (en) | Apparatus and method for improving an audio signal in the spectral domain | |
| JP2003169119A (en) | Speaker assembly for mobile phone | |
| WO2018139884A1 (en) | Method for processing vr audio and corresponding equipment | |
| WO2014148844A1 (en) | Terminal device and audio signal output method thereof | |
| KR20140055932A (en) | Apparatus and method for keeping output loudness and quality of sound among differernt equalizer modes | |
| EP3748635B1 (en) | Acoustic device and acoustic processing method | |
| WO2014148845A1 (en) | Audio signal size control method and device | |
| WO2019084886A1 (en) | Audio denoising method and denoising device, mobile terminal and readable storage medium | |
| CN113286244A (en) | Microphone anomaly detection method and device | |
| WO2019019329A1 (en) | Method and system for optimizing network performance and computer readable storage medium | |
| GB2500251A (en) | Active noise cancellation system with wind noise reduction | |
| US9124985B2 (en) | Hearing aid and method for automatically controlling directivity | |
| WO2015180430A1 (en) | Voice control method and system |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 14804561 Country of ref document: EP Kind code of ref document: A1 |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 32PN | Ep: public notification in the ep bulletin as address of the adressee cannot be established |
Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205 DATED 07/04/2016) |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 14804561 Country of ref document: EP Kind code of ref document: A1 |