EP3685378A1 - Signal processor and method for providing a processed audio signal reducing noise and reverberation - Google Patents
Signal processor and method for providing a processed audio signal reducing noise and reverberationInfo
- Publication number
- EP3685378A1 EP3685378A1 EP18769221.5A EP18769221A EP3685378A1 EP 3685378 A1 EP3685378 A1 EP 3685378A1 EP 18769221 A EP18769221 A EP 18769221A EP 3685378 A1 EP3685378 A1 EP 3685378A1
- Authority
- EP
- European Patent Office
- Prior art keywords
- signal
- noise
- reduced
- reverberation
- signal processor
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Granted
Links
Classifications
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
- G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
- G10L21/0208—Noise filtering
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
- G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
- G10L21/0208—Noise filtering
- G10L21/0216—Noise filtering characterised by the method used for estimating noise
- G10L21/0232—Processing in the frequency domain
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
- G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
- G10L21/0208—Noise filtering
- G10L21/0264—Noise filtering characterised by the type of parameter measurement, e.g. correlation techniques, zero crossing techniques or predictive techniques
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
- G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
- G10L21/0208—Noise filtering
- G10L2021/02082—Noise filtering the noise being echo, reverberation of the speech
Definitions
- Embodiments according to the invention are related to a signal processor for providing a processed audio signal.
- Further embodiments according to the invention are related to a method for providing a processed audio signal. Further embodiments according to the invention are related to a computer program for performing said methods.
- Embodiments according to the invention are related to a method and apparatus for online dereverberation and noise reduction (for example, using a parallel structure) with reduction control.
- Embodiments according to the invention are related to linear prediction based online dereverberation and noise reduction using alternating Kalman filters.
- Embodiments according to the invention relate to a signal processor, a method and a computer program for noise reduction and reverberation reduction.
- Audio signal processing, speech communication and audio transmission are continuously developing technical fields. However, when handling audio signals, it is often found that noise and reverberation degrade the audio quality.
- the speech quality and intelligibility is typically degraded due to high levels of reverberation and noise compared to the desired speech level.
- Advantages of methods based on the MAR model are that they are valid for multiple sources, they directly estimate a dereverberation filter of finite length, the required filters are relatively short, and they are suitable as pre-processing techniques for beamforming algorithms.
- a great challenge of the MAR signal model is the integration of additive noise, which has to be removed in advance [30], [32] without destroying the relations between neighboring time-frames of the reverberant signal.
- MAR signal model a generalized framework for the multichannel linear prediction methods called blind impulse response shortening was presented, which aims at shortening the reverberant tail in each microphone and results in the same number of output as input channels, while preserving the inter-microphone correlation of the desired signal.
- An embodiment according to the invention creates a signal processor for providing a processed audio signal (for example, a noise-reduced and reverberation-reduced audio signal, which may be a single-channel audio signal or a multi-channel audio signal) (or generally speaking, one or more processed audio signals) on the basis of an input audio signal (for example, a single-channel or a multi-channel input audio signal) (or generally speaking, on the basis of one or more input audio signals).
- a processed audio signal for example, a noise-reduced and reverberation-reduced audio signal, which may be a single-channel audio signal or a multi-channel audio signal
- an input audio signal for example, a single-channel or a multi-channel input audio signal
- the signal processor is configured to estimate coefficients of an (for example, multi-channel) autoregressive reverberation model (for example, AR coefficients or MAR coefficients) using the input audio signal (for example, the noisy and reverberant input audio signal or multiple noisy and reverberant input audio signals, or directly an observed signal y(n) which may, for example, originate from one or more microphones) (or, generally speaking, using one or more input audio signals) and (one or more) delayed noise-reduced reverberant signals obtained using a noise reduction (or a noise reduction stage).
- the delayed noise-reduced reverberant signal may comprise (one or more) past noise-reduced reverberant signals which may be represented by x( 7).
- the estimation of the coefficients may be performed by an AR coefficient estimation stage or by an MAR coefficient estimation stage of the signal processor.
- the signal processor is configured to provide a noise-reduced reverberant signal (for example, of a current frame) (or, generally speaking, one or more noise- reduced reverberant signals) using the input audio signal (which may, for example, be a noisy and reverberant input audio signal or which may, for example, be the noisy observed signal y(n) which may originate from one or more microphones) and the estimated coefficients of the autoregressive reverberation model (which may be a multi- channel autoregressive reverberation model) (and wherein the estimated coefficients may, for example, be associated with the current frame and may, for example, be called "MAR coefficients").
- the input audio signal which may, for example, be a noisy and reverberant input audio signal or which may, for example, be the noisy observed signal y(n) which may originate from one or more microphones
- the estimated coefficients of the autoregressive reverberation model which may be a multi- channel autoregressive re
- the part of the signal processor configured to provide the noise- reduced reverberant signal may be considered as a "noise reduction stage".
- the audio signal processor is configured to provide a noise-reduced and reverberation-reduced output signal (or, generally speaking, one or more noise-reduced and reverberation-reduced output signals) using the noise-reduced (reverberant) signal (or, generally speaking, one or more noise-reduced, reverberant signals) and the estimated coefficients of the autoregressive reverberation model (or multi-channel autoregressive reverberation model). This may, for example, be performed using a reverberation estimation and a signal subtraction.
- This embodiment according to the invention is based on the finding that it is possible to overcome a causality problem, which is found in some conventional solutions, by estimating the coefficients of the autoregressive reverberation model associated with a certain frame on the basis of a delayed and noise reduced reverberant signal which may be associated with one or more preceding frames, and that it is possible to provide the noise reduced reverberant signal of the current frame using the input audio signal and the estimated coefficients of the autoregressive reverberation model associated with the current frame and obtained on the basis of noise-reduced (and typically reverberant) signals (for example, provided by the noise reduction stage) associated with one or more preceding frames.
- the computational complexity can be kept reasonably small, since the estimation of the coefficients of the autoregressive reverberation model and the estimation of the noise-reduced reverberant signal can be performed separately and alternatingly.
- the separate estimation of the coefficients of the autoregressive reverberation model and of the noise-reduced reverberant signal can be performed more efficiently than a joint estimation of coefficients of an autoregressive reverberation model and of a noise-reduced reverberant signal, and also more efficiently than a joint (one-step) estimation of a noise-reduced and reverberation-reduced audio signal.
- the signal processor is configured to estimate coefficients of a multi-channel autoregressive reverberation model. It has been found that the concept described herein is well-suited for a handling of multi-channel signals and brings along particular improvements of the complexity for such multi-channel signals.
- the signal processor is configured to use estimated coefficients of the autoregressive reverberation model associated with a currently processed portion (for example, a time-frame having a frame index n) of the input audio signal in order to produce the noise-reduced reverberant signal associated with the currently processed portion (for example, a time-frame having frame index n) of the input audio signal.
- the provision of the noise-reduced reverberant signal associated with the currently processed portion may rely on the previous estimation of the coefficients of the autoregressive reverberation model associated with the currently processed portion of the input audio signal, or the estimation of the coefficients of the autoregressive reverberation model associated with a currently processed portion (or frame) may precede the provision of the noise-reduced reverberant signal associated with the currently processed portion (or frame).
- the estimation of the coefficients of the autoregressive reverberation model may be performed first (for example, using a past noise reduced but reverberant signal) and the provision of the noise-reduced reverberant signal associated with the currently processed frame may be performed then. It has been found that such an order of the processing results in particularly good results, while a reverse order will typically not perform quite as good.
- the signal processor is configured to use one or more delayed noise-reduced reverberant signals (or, alternatively, a noise-reduced reverberant signal) associated with (or based on) a previously processed portion (for example, a frame having frame index n-1 ) of the input audio signal (for example, an input signal y(n)) for an estimation of coefficients of the autoregressive reverberation model associated with the currently processed portion (for example, having a frame index n) of the input audio signal.
- a previously processed portion for example, a frame having frame index n-1
- the input audio signal for example, an input signal y(n)
- a causality problem can be avoided, since the provision of the noise-reduced reverberant signal associated with the previously processed frame can typically be provided before the estimation of the coefficients of the autoregressive reverberation model associated with the currently processed portion (or frame) of the input audio signal. Also, it has been found that the usage of a noise reduced reverberant signal associated with a previously processed portion of the input audio signal results in a sufficiently good estimation of the coefficients of the autoregressive reverberation model.
- the signal processor is configured to alternatingly provide estimated coefficients of the autoregressive reverberation model (or multi-channel autoregressive reverberation model) and noise-reduced reverberant signal portions. Moreover, the signal processor is configured to use estimated coefficients (or, alternatively, previously estimated coefficients) of the (preferably multi-channel) autoregressive reverberation model for the provision of the noise-reduced reverberant signal portions. Moreover, the signal processor is configured to use one or more delayed noise-reduced reverberant signals (or, alternatively, previously provided noise reduced reverberant signal portions) for the estimation of coefficients of the multi-channel autoregressive reverberation model.
- the computational complexity can be kept low and results can still be obtained with little delay. Also, computational instabilities, which could be caused by a joint estimation of coefficients of the multi-channel autoregressive reverberation model and noise reduced reverberant signal portions can be avoided.
- the signal processor may be configured to apply an algorithm minimizing a cost function (for example, a Kalman filter, a recursive least squares filter or a normalized least mean squares (NLMS) filter) in order to estimate the coefficients of the (preferably multi-channel) autoregressive reverberation model.
- a cost function for example, a Kalman filter, a recursive least squares filter or a normalized least mean squares (NLMS) filter
- the cost function may, for example be defined as shown in equation (15), and the minimization may, for example, fulfill the functionality as shown in equation (17) or minimize the trace of an error matrix, as shown in equation (19).
- the Minimization of the cost function may, for example, follow equations (20) to (25).
- the minimization of the cost function may also use steps 4 to 6 of Algorithm 1.
- the cost function used for the estimation of the coefficients of the autoregressive reverberation model is an expectation value for a mean squared error of the coefficients of the autoregressive reverberation model, for example, as shown in equation (19). Accordingly, coefficients of the autoregressive reverberation model which are expected to fit well an acoustic environment causing the reverberation can be achieved. It should be noted that expected statistical properties of the MAR coefficient noise and of the noisy dereverberated signals (state and observation noises), for example, be estimated in a separate, preparatory step (for example, using one or more of equations (26) to (29).
- the signal processor may be configured to apply the algorithm for the minimization of the cost function in order to estimate the coefficients of the (preferably multi-channel) autoregressive reverberation model under the assumption that the noise-reduced reverberant signal is fixed (for example, not affected by the coefficients of the autoregressive reverberation model associated with the currently processed portion of the input audio signal).
- the algorithm of equations (20) to (25) makes such an assumption.
- the signal processor is configured to apply an algorithm for a minimization of a cost function (for example, a Kalman filter or a recursive least squares filter or a NLMS filter) in order to estimate the noise-reduced reverberant signal.
- a cost function for example, a Kalman filter or a recursive least squares filter or a NLMS filter
- the cost function may, for example be defined as shown in equation (16), and the minimization may, for example, fulfill the functionality as shown in equation (18) or minimize the trace of an error matrix, as shown in equation (30).
- the minimization of the cost function may, for example, follow equations (31 ) to (36).
- the signal processor is configured to apply an algorithm for a minimization of a cost function (for example, a Kalman filter , a recursive least squares filter or a NLMS filter) in order to estimate the noise-reduced reverberant signal.
- a cost function for example, a Kalman filter , a recursive least squares filter or a NLMS filter. It has been found that the usage of such an algorithm for a minimization of a cost function is also very efficient for the determination of the noise-reduced reverberant signal, for example, if statistical properties of the noise are known or estimated.
- the computational complexity can be substantially improved if similar algorithms (for example, algorithms minimizing a cost function) are used both for the estimation of the coefficients of the autoregressive reverberation model and for the estimation of the noise-reduced reverberant signal.
- similar algorithms for example, algorithms minimizing a cost function
- the algorithm according to equations (31 ) to (36) may be used, wherein parameters to be used in said algorithm may be determined according to one or more of equations (37) to (42).
- the functionality may be performed using steps 7 to 9 of Algorithm 1 .
- the cost function used for the estimation of the (optionally noise-reduced) reverberant signal is an expectation value for a mean-squared error of the (optionally noise-reduced) reverberant signal. It has been found that such a cost function (for example, according to equation (16) or according to equation (30)) provides for good results and can be evaluated using reasonable computational effort. Moreover, it should be noted that the estimation of the mean squared error of the noise-reduced reverberant signal is possible, for example, if information (or assumption) regarding statistical characteristics of the noise (for example, the noise covariance matrix) and possibly also regarding the desired signal (for example, the desired speech covariance matrix) are available.
- the signal processor is configured to apply the algorithm for the minimization of the cost function in order to estimate the (optionally noise-reduced) reverberant signal under the assumption that the coefficients of the autoregressive reverberation model are fixed (for example, not affected by the noise-reduced reverberant signal associated with the currently processed portion of the input audio signal).
- the assumption allows for an alternating procedure in which the noise- reduced reverberant signal and the coefficients of the autoregressive reverberation model are estimated in a separated manner (for example, by alternatingly performing steps 4 to 6 and steps 7 to 9 of Algorithm 1 ).
- the signal processor is configured to determine a reverberation component on the basis of estimated coefficients of the (preferably multichannel) autoregressive reverberation model and on the basis of one or more delayed noise-reduced reverberant signals (or, alternatively, on the basis of the noise-reduced reverberant signal) associated with a previously processed portion (for example, a frame) of the input audio signal (for example, by filtering the noise-reduced reverberant signal using the estimated coefficients of the autoregressive reverberation model).
- the signal processor is preferably configured to (at least partially) cancel (for example, subtract) the reverberation component from the noise-reduced reverberant signal associated with a currently processed portion (for example, a frame) of the input audio
- the signal processor is configured to estimate a statistic (for example, a covariance) (or a statistical property) of a noise component of the input audio signal.
- a statistic of the noise component of the input audio signal may, for example, be useful in the estimation (or provision) of a noise-reduced reverberant signal.
- an estimation (or determination) of a statistic of the noise component of the input audio signal can facilitate a formulation of a cost function because the statistic of the noise component of the input audio signal can be used as a part of said cost function.
- the signal processor is configured to estimate a statistic (for example, a covariance) (or a statistical property) of a noise component of the input audio signal during a non-speech period (wherein, for example, the non-speech period is detected using a speech detector).
- a statistic for example, a covariance
- the noise which is present during non-speech periods is typically also present during the speech periods without too many changes. Accordingly, it is possible to efficiently obtain the statistics of the noise component, which are useable for the provision of the noise-reduced reverberant signal.
- the signal processor is configured to estimate the coefficients of the (preferably multi-channel) autoregressive reverberation modeled using a Kalman filter. It has been found that such a Kalman filter allows for an efficient computation and is well-adapted to the requirements of the signal processing task. For example, the implementation according to equations (20) to (25) can be used.
- the signal processor is configured to estimate the coefficients of the (preferably multi-channel) autoregressive reverberation model on the basis of an estimated error matrix of a vector of coefficients of the (preferably multi-channel) autoregressive reverberation model (for example, associated with a previously processed portion of the audio signal), on the basis of an estimated covariance of an uncertainty noise of the vector of a coefficient of the (preferably multi-channel) autoregressive reverberation model (for example, as given in equation (26)), on the basis of a previous vector of (estimated) coefficients of the (preferably multi-channel) autoregressive reverberation model (for example, associated with a previously processed portion or version of the input audio signal), on the basis of one or more delayed noise-reduced reverberant signals delayed noise-reduced reverberant signals (for example, (past) noise- reduced reverberant signals, represented by X(n), for example associated with previous portions or frames of the input audio signal),
- the signal processor is configured to estimate the noise- reduced reverberant signal using a Kalman filter. It has been found that usage of such a Kalman filter (which may implement the functionality as given in equations 31 to 36) is also advantageous for the estimation of the noise-reduced reverberant signal. Also, using a Kalman filter both for the estimation of the coefficient of the autoregressive reverberation model and for the estimation of the noise-reduced reverberant signal can provide good results.
- the signal processor is configured to estimate the noise- reduced reverberant signal on the basis of an estimated error matrix of the noise-reduced reverberant signal (for example, associated with a previously-processed portion or frame of the input audio signal, for example), on the basis of an estimated covariance of a desired speech signal (for example, associated with a currently processed portion or frame of the input audio signal, for example, as given in equations 37 to 42), on the basis of one or more previous estimates of the noise-reduced reverberant signal (for example, associated with one or more previously processed portions or frames of the input audio signal), on the basis of a plurality of coefficients of the (preferably multi-channel) autoregressive reverberation model (for example, associated with the currently processed portion or frame of the input audio signal, for example defining a matrix F(n)), on the basis of an estimated noise covariance associated with the input audio signal, and on the basis of the input audio signal.
- the signal processor is configured to obtain an estimated covariance associated with noisy but reverberation-reduced (or non-reverberant) signal components of the input audio signal on the basis of a weighted combination (for example, according to equation 28) of a recursive covariance estimate determined recursively using previous estimates of noisy but reverberation-reduced (or non- reverberant) signal components of the input audio signal (for example, associated with previously processed portions or frames of the input audio signal, for example according to equation 29) and of an outer product of an (for example, intermediate) estimate of noisy but reverberation-reduced (or non-reverberant) signal components of the input audio signal (for example, associated with a currently processed portion of the input audio signal).
- a weighted combination for example, according to equation 28
- a recursive covariance estimate determined recursively using previous estimates of noisy but reverberation-reduced (or non- reverberant) signal components of the input audio
- the intermediate estimate of the noisy but reverberation-reduced signal components may be obtained as an innovation in a Kalman filtering process (for example, according to equation (22)).
- the intermediate estimate may be a prediction using predicted coefficients (for example, as determined by equation (21 )).
- the recursive covariance estimate of the desired signal plus noise is based on an estimation of the noisy but reverberation-reduced (or non- reverberant) signal components of the input audio signal computed using final estimate coefficients of the (preferably multi-channel) autoregressive reverberation model and using a final estimate of the noise-reduced reverberant signal (for example, according to equation (29) in combination with the definition of u(n)).
- the signal processor is configured to obtain the outer product of the noisy but reverberation- reduced signal components of the input audio signal on the basis of an intermediate estimate (for example, a prediction) of the coefficients of the (preferably multi-channel) autoregressive reverberation model (for example, in a Kalman filtering process) (for example, in order to obtain the covariance estimate)(for example obtained according to equation (21 )).
- an intermediate estimate for example, a prediction
- the coefficients of the (preferably multi-channel) autoregressive reverberation model for example, in a Kalman filtering process
- the covariance estimate for example obtained according to equation (21 )
- the signal processor is configured to obtain an estimated covariance associated with a noise-reduced and reverberation-reduced (or non- reverberant) signal component of the input audio signal on the basis of a weighted combination (for example, according to equation (37)) of a recursive covariance estimate determined recursively using previous estimates of a noise-reduced and reverberation- reduced signal components of the input audio signal (for example, associated with previously processed portions or frames of the input audio signal) (which may, for example, be considered as a recursive a-posteriori maximum likelihood estimate) and of an a-priori estimate of the covariance which is based on a currently processed portion of the input audio signal (and obtained, for example, in accordance with equation (41 )).
- a weighted combination for example, according to equation (37)
- a recursive covariance estimate determined recursively using previous estimates of a noise-reduced and reverberation- reduced signal components
- the signal processor is configured to obtain the recursive covariance estimate based on an estimation of the noise-reduced and the reverberation- reduced (or non-reverberant) signal components of the input audio signal computed using final estimated coefficients of the (preferably multi-channel) autoregressive reverberation model and using a final estimate of the noise-reduced reverberant (output) signal (for example, using equation (38)).
- the signal processor is configured to obtain the a-priori estimate of the covariance using a Wiener filtering of the input signal (as shown, for example, in equation (41 )), wherein a Wiener filtering operation is determined in dependence on the covariance information regarding the input audio signal, in dependence on covariance information regarding a reverberation component of the input audio signal and in dependence on covariance information regarding a noise component of the input audio signal (as shown, for example, in equation (42)). It has been found that these concepts are helpful in efficient computation of the estimated covariance associated with the noise-reduced and reverberation-reduced signal component.
- Another embodiment according to the invention creates a method for providing a processed audio signal (for example, a noise-reduced and reverberation-reduced audio signal, which may be a single-channel audio signal or a multi-channel audio signal) on the basis of an input audio signal (for example, a single-channel or multi-channel input audio signal).
- a processed audio signal for example, a noise-reduced and reverberation-reduced audio signal, which may be a single-channel audio signal or a multi-channel audio signal
- the method comprises estimating coefficients of a (preferably, but not necessarily, multi-channel) autoregressive reverberation model (for example, AR coefficients or MAR coefficients) using the ⁇ typically noisy and reverberant) input audio signal (or input audio signals) (for example, directly from the observed signal y(n)) and delayed (or past) noise-reduced reverberant signals obtained using a noise reduction (noise reduction stage) (for example, past noise-reduced reverberant signals x(f?)).
- This functionality may, for example, be performed by the AR coefficient estimation stage.
- the method comprises providing a noise-reduced reverberant signal (for example, of a current frame) using the (typically noisy and reverberant) input audio signal (for example, the noisy observed signal y(n)) and the estimated coefficients of the (preferably multi-channel) autoregressive reverberation model (for example, associated with the current frame).
- the estimated coefficients of the autoregressive reverberation model may, for example, be "MAR coefficients".
- the functionality of providing the noise-reduced reverberant signal may, for example, be performed by a noise reduction stage.
- the method further comprises deriving a noise-reduced and reverberation-reduced output signal using the noise-reduced reverberant signal and the estimated coefficients of the (preferably multi-channel) autoregressive reverberation model.
- Another embodiment according to the invention creates a computer program for performing the method as described herein when the computer program runs on a computer.
- the nput audio signal 1 10 can be a single-channel audio signal but is preferably a multi-channel audio signal.
- the processed audio signal 1 12 can be a single-channel audio signal but is preferably a multi-channel audio signal.
- the signal processor 100 may, for example, comprise a coefficient estimation block or coefficient estimation unit 120, which is configured to estimate coefficients 124 of an autoregressive reverberation model (for example, AR coefficients or MAR coefficients of a multi-channel autoregressive reverberation model) using the single-channel or multichannel input audio signal 1 10 and a delayed noise-reduced reverberant signal 122.
- an autoregressive reverberation model for example, AR coefficients or MAR coefficients of a multi-channel autoregressive reverberation model
- the estimation of the coefficients of the autoregressive reverberation model 120 may receive the input audio signal 1 10 and the delayed noise-reduced reverberant signal 122.
- the signal processor 100 also comprises a noise reduction unit or noise reduction block 130 which receives the input audio signal 1 10 and which provides a noise-reduced (but typically reverberant or non-reverberation-reduced) signal 132.
- the noise reduction unit or noise reduction block 130 is configured to provide a noise-reduced (but typically reverberant) signal using the (typically noisy and reverberant) input audio signal 1 10 and the estimated coefficients 124 of the autoregressive reverberation model which are provided by the estimation block or estimation unit 120.
- the noise reduction 130 may, for example, use coefficients 124 of the autoregressive reverberation model which have been obtained on the basis of a previously determined noise-reduced reverberant signal 132 (possibly in combination with the input audio signal 1 10).
- the apparatus 100 optionally comprises a delay block or delay unit 140, which may be configured to obtain the noise-reduced reverberant signal 132 provided by the noise reduction unit or noise reduction block 130 to provide, as an output, a delayed version 122 thereof. Accordingly, the estimation 120 of the coefficients of the autoregressive reverberation model can operate on a previously obtained (derived) noise-reduced reverberant signal (which is provided or derived by the noise reduction block 130) and the input audio signal 1 10.
- a delay block or delay unit 140 may be configured to obtain the noise-reduced reverberant signal 132 provided by the noise reduction unit or noise reduction block 130 to provide, as an output, a delayed version 122 thereof. Accordingly, the estimation 120 of the coefficients of the autoregressive reverberation model can operate on a previously obtained (derived) noise-reduced reverberant signal (which is provided or derived by the noise reduction block 130) and the input audio signal 1 10.
- the apparatus 100 also comprises a block or unit 150 for the derivation of a noise- reduced and reverberation-reduced output signal, which may serve as the processed audio signal 1 12.
- the block or unit 150 preferably receives the noise-reduced reverberant signal 132 from the noise reduction block or noise reduction unit 130 and the coefficients 124 of the autoregressive reverberation model provided by the estimation block or estimation unit 120.
- the block or unit 150 may, for example, remove or reduce reverberation from the noise-reduced reverberant signal 132.
- an appropriate filtering in combination with a cancellation operation (for example, in a spectral domain) may be used for this purpose, wherein the coefficients 124 of the autoregressive reverberation model may determine the filtering (which is used to estimate the reverberation).
- the separation of functionalities into blocks or units can be considered as an efficient but arbitrary choice.
- the functionalities described herein could also be distributed differently to a hardware apparatus as long as the fundamental functionality is maintained.
- the blocks or units could be software blocks or software units which reuse the same hardware (like, for example, a microprocessor).
- the separation between the noise reduction functionality (noise reduction block or noise reduction unit 130) and the estimation of the coefficients of the autoregressive reverberation model (estimation block or estimation unit 120) provides for a reasonably small computational complexity and still allows for obtaining a sufficiently good audio quality. Even though, theoretically, it would be best to estimate the noise-reduced and reverberation-reduced output signal using a joint cost function, it has been found that separately performing the noise reduction and the estimation of the coefficients of the autoregressive reverberation model using separate cost functions can still provide reasonably good results, while complexity can be reduced and stability problems can be avoided.
- the noise-reduced reverberant signal 132 serves as a very good intermediate quality, since the noise-reduced and reverberation-reduced output signal (i.e., the processed audio signal 1 12) can be derived from the noise-reduced (but reverberant or non- reverberation-reduced) signal 132 with little effort provided that the coefficients 124 of the autoregressive reverberation model are known.
- the following embodiments of the invention are in the field of acoustic field processing, for example to remove reverberation noise from one or multiple microphones.
- the speech quality and intelligibility as well as the performance of speech recognizers is typically degraded due to high levels of reverberation and noise compared to the desired speech level.
- Dereverberation methods based on an autoregressive (AR) model per frequency band in the short-time Fourier transform (STFT) domain have been shown to perform superior to other reverberation models. Dereverberation methods based on this model typically solve the problem using approaches related to linear prediction. Furthermore, the general multichannel autoregressive (MAR) model is valid for multiple sources and can be formulated such that it provides the same number of channels at the output as at the input. Since the resulting enhancement process, which is a linear filter per frequency band across multiple STFT frames, does not change the spatial correlation of the desired signal, the enhancement is suitable as preprocessing for further array processing techniques.
- AR autoregressive
- STFT short-time Fourier transform
- the problem can be typically be solved by first performing a noise reduction step, followed by linear prediction-based methods to estimate the MAR coefficients (also known as room regression coefficients) and then filtering the signal.
- MAR coefficients also known as room regression coefficients
- a noise reduction stage 202 tries to remove the noise from the observed signals y(n) , and in a second step 203 the AR coefficients c(n) are estimated from the output signals of the first stage X(fj). It has been found that this structure is suboptimal for two reasons: 1 ) The MAR parameter estimation stage 203 assumes that the estimated signal x(n) is noise-free, which is often not possible in practice.
- Fig. 2 shows a block schematic diagram of a conventional structure for MAR coefficient estimation in a noisy environment.
- the apparatus 200 comprises a noise statistics estimation 201 , a noise reduction 202, an AR coefficient estimation 203 and a reverberation estimation 204.
- blocks 201 to 204 are blocks of the conventional sequential noise reduction and the reverberation system.
- FIG. 3 shows a block schematic diagram of embodiment 2 according to the present invention.
- Fig. 4 shows a block schematic diagram of embodiment 3 according to the present invention.
- Fig. 5 shows a block schematic diagram of embodiment 4 according to the present invention.
- the apparatus 300 also comprises an autoregressive coefficient estimation 302 (AR coefficient estimation) which is configured to receive the input audio signal 301 and a delayed version (or past version) of the noise-reduced (but typically reverberant) signal 303a provided by the noise reduction 303. Moreover, the autoregressive coefficient estimation 302 is configured to provide the coefficients 302a of the autoregressive reverberation model.
- AR coefficient estimation an autoregressive coefficient estimation
- the apparatus 300 optionally comprises a delayer 320 which is configured to derive the delayed version 320a from the noise-reduced (but typically reverberant) signal 303a provided by the noise reduction 303.
- the apparatus 300 also comprises a reverberation estimation 304, which is configured to receive the delayed version 320a of the noise-reduced (but typically reverberant) signal 303a provided by the noise reduction 303. Moreover, the reverberation estimation 304 also receives the coefficients 302a of the autoregressive reverberation model from the autoregressive coefficient estimation 302. The reverberation estimation 304 provides an estimated reverberation signal 304a.
- the apparatus 300 also comprises a signal subtractor 330 which is configured to remove (or subtract) the estimated reverberation signal 304a from the noise-reduced (but typically reverberant) signal 303a provided by the noise reduction 303, to thereby obtain the processed audio signal 312, which is typically noise-reduced and reverberation-reduced.
- a signal subtractor 330 which is configured to remove (or subtract) the estimated reverberation signal 304a from the noise-reduced (but typically reverberant) signal 303a provided by the noise reduction 303, to thereby obtain the processed audio signal 312, which is typically noise-reduced and reverberation-reduced.
- the autoregressive coefficient estimation 302 uses both the input signal 310 and the noise-reduced (but typically reverberant) output signal 303a of the noise reduction 303 (or, more precisely, a delayed version 320a thereof).
- the autoregressive coefficient estimation 302 can be performed separately from the noise reduction 303, wherein the noise reduction 303 can nevertheless take benefit of the coefficients 302a of the autoregressive reverberation model, and wherein the autoregressive coefficient estimation 302 can nevertheless take benefit of the noise-reduced signal 303a provided by the noise reduction 303.
- the reverberation can finally be removed from the noise-reduced (but typically reverberant) signal 303a provided by the noise reduction 303.
- the apparatus or signal processor 500 according to Fig. 5 is similar to the apparatus or signal processor 400 according to Fig. 4, such that reference is made to the above explanations and such that equal components will not be described again.
- the apparatus 500 also comprises a reverberation shaping 305 which receives the reverberation signal 304a provided by the reverberation estimation.
- the reverberation shaping 305 provides a shaped reverberation signal 305a.
- the reverberation signal 304a is subtracted from the sum of the scaled noise reduced signal 303b and the scaled input signal 410a. accordingly, an intermediate signal 520 is obtained.
- a scaled version 305b of the shaped reverberation signal 305a is added to the intermediate signal 520 in order to obtain an output signal 512.
- the apparatus 500 allows to adjust characteristics of the output signal 512.
- the original reverberation can be removed (at least to a large degree), for example by subtracting the (estimated) reverberation signal 304a from the sum of signals 303b, 410a. Accordingly, a modified (shaped) reverberation signal 305b can be added (for example after an optional scaling), to thereby obtain the output signal 512. Accordingly, the output signal can be obtained with a shaped reverberation and with an adjustable degree of noise reduction.
- the parallel structure shown in Fig. 3 allows for an easy and effective way to control the amount of reverberation and noise reduction.
- Such a control can be desired in speech communication scenarios to keep e.g., some residual noise and reverberation for perceptual reasons or to mask artifacts produced by the reduction algorithm.
- We define the (desired) new output signal z(n) s(n)+#r(n)+A,v(A7), where ⁇ ⁇ and are the control parameters for the residual reverberation and noise.
- an optional processing of the reverberation signal f(f?) can be inserted as shown in Fig. 4 in Block
- the output signal with reverberation shaping is then computed by - ⁇ ⁇ ⁇ (")+ (1 - ⁇ ⁇ ) ⁇ )- ⁇ )+ ⁇ ⁇ ), where r s (n) is the shaped reverberation signal by Block 305.
- the reverberation shaping can be performed for example by an equalizer or compressor / expander commonly used in audio and music production.
- Multi-channel linear prediction based dereverberation in the short-time Fourier transform (STFT) domain has been shown to be highly effective.
- STFT short-time Fourier transform
- MAR multi-channel autoregressive
- the proposed method is evaluated using simulated and measured acoustic impulse responses and compared to a method based on the same signal model.
- a method (and concept) to control the amount of reverberation and noise reduction independently is described.
- Embodiments according to the invention can be used for a dereverberation.
- Embodiments according to the invention use a multi-channel linear prediction and an autoregressive model.
- Embodiments according to the invention use a Kalman filter, preferably in combination with an alternating minimization.
- a method (and concept) based on the MAR reverberation model is proposed to reduce reverberation and noise using an online algorithm.
- the proposed solution outperforms the noise-free solution presented in [3] where the MAR coefficients are modeled by a time-varying first-order Markov model. To obtain the desired dereverberated speech signals, it is possible to estimate the MAR coefficients and the noise-free reverberant speech signal.
- the proposed solution has several advantages to conventional solutions: Firstly in contrast to the sequential signal and autoregressive (AR) parameter estimation methods used for noise reductions presented in [8] and [17], a parallel estimation structure as an alternating minimization algorithm using, for example, two interactive Kalman filters to estimate the MAR coefficients and the noise-free reverberant signals is proposed. This parallel structure allows a fully causal estimation chain as opposed to a sequential structure, where the noise reduction stage would use outdated MAR coefficients.
- AR autoregressive
- subsection 2 the signal models for the reverberant signal, the noisy observation and the MAR coefficients are presented and the problem is formulated.
- subsection 3 two alternating Kalman filters are derived as part of an alternating minimization problem to estimate the MAR coefficients and the noise-free signals.
- An optional method to control the reverberation and noise reduction is presented in subsection 4.
- subsection 5 the proposed method and concept is evaluated and compared to state-of-the-art methods.
- estimated quantities may optionally take the place of ideal quantities.
- the filter may be time-varying, wherein it is assumed that a previous set of filter coefficients is scaled by a matrix A and affected by a "process noise" w(n).
- MAP-EM In the method proposed in [31], the MAR coefficients are estimated using a Bayesian approach based on MAP estimation and the noise-free desired signal is then estimated using an EM algorithm. The algorithm is online, but the EM procedure requires about 20 iterations per frame to converge.
- the measures for the noisy reverberant input signal are indicated as light grey dashed line, and the SRMR of the target signal, i. e. the early speech, is indicated as dark grey dash-dotted line.
- the CD is larger than for the input signal, which indicates an overall quality deterioration, whereas PESQ, SIR and SRMR still improve over the input, i. e. reverberation and noise are reduced.
- the performance in terms of all measures increases by increasing the number of microphones.
- Embodiments according to the invention can optionally comprise one or more of the following features:
- the MAR coefficients are estimated using the noisy reverberant input signals and delayed estimated reverberant output signals from the noise reduction stage.
- the noise reduction stage receives current MAR coefficient estimates in each frame (optional).
- an audio encoder apparatus for providing an encoded representation of an input audio signal
- an audio decoder apparatus for providing a decoded representation of an audio signal on the basis of an encoded representation.
- any of the features described herein can be used in the context of an audio encoder and in the context of an audio decoder.
- features and functionalities disclosed herein relating to a method can also be used in an apparatus (configured to perform such a method or functionality).
- any of the features and functionalities disclosed herein with respect to an apparatus can also be used in a corresponding method.
- the methods disclosed herein can be supplemented by any of the features and functionalities described with respect to the apparatuses and vice versa.
- any of the features and functionalities described herein can be implemented in hardware and software (or using hardware and/or software), or even a combination of hardware and software, as will be described in the section "Implementation Alternatives".
- processing described herein may be performed, for example (but not necessarily) per frequency band or per frequency bin or for different frequency regions. It should be noted that aspects of the invention relate to a method and apparatus for online dereverberation and noise reduction with reduction control.
- Embodiments according to the invention create a novel parallel structure for joint dereverberation and noise reduction.
- the reverberant signal is modelled, for example, using a narrowband multichannel autoregressive reverberation model with time-varying coefficients, which account for non-stationary acoustic environments.
- embodiments according to the invention estimate the noise-free reverberant signal and the autoregressive room coefficients in parallel, such that assumptions on stationary room coefficients are not required.
- a method to independently control the reduction level of noise and reverberation is proposed.
- Fig. 14 shows a flow chart of a method 1400 according to an embodiment of the present invention.
- the method 1400 for providing a processed audio signal on the basis of an input audio signal comprises estimating 1410 coefficients of an autoregressive reverberation model using the input audio signal and a delayed noise-reduced reverberant signal obtained using a noise reduction stage.
- the method also comprises providing 1420 a noise-reduced reverberant signal using the input audio signal and the estimated coefficients of the autoregressive reverberation model.
- the method also comprises deriving 1430 a noise-reduced and reverberation-reduced output signal using the noise-reduced reverberant signal and the estimated coefficients of the autoregressive reverberation model.
- the method 1400 can optionally be supplemented by any of the features, functionalities and details describer herein, both individually and in combination.
- aspects have been described in the context of an apparatus, it is clear that these aspects also represent a description of the corresponding method, where a block or device corresponds to a method step or a feature of a method step. Analogously, aspects described in the context of a method step also represent a description of a corresponding block or item or feature of a corresponding apparatus.
- Some or all of the method steps may be executed by (or using) a hardware apparatus, like for example, a microprocessor, a programmable computer or an electronic circuit. In some embodiments, one or more of the most important method steps may be executed by such an apparatus.
- embodiments of the invention can be implemented in hardware or in software.
- the implementation can be performed using a digital storage medium, for example a floppy disk, a DVD, a Blu-Ray, a CD, a ROM, a PROM, an EPROM, an EEPROM or a FLASH memory, having electronically readable control signals stored thereon, which cooperate (or are capable of cooperating) with a programmable computer system such that the respective method is performed. Therefore, the digital storage medium may be computer readable.
- Some embodiments according to the invention comprise a data carrier having electronically readable control signals, which are capable of cooperating with a programmable computer system, such that one of the methods described herein is performed.
- embodiments of the present invention can be implemented as a computer program product with a program code, the program code being operative for performing one of the methods when the computer program product runs on a computer.
- the program code may for example be stored on a machine readable carrier.
- inventions comprise the computer program for performing one of the methods described herein, stored on a machine readable carrier.
- an embodiment of the inventive method is, therefore, a computer program having a program code for performing one of the methods described herein, when the computer program runs on a computer.
- a further embodiment of the inventive methods is, therefore, a data carrier (or a digital storage medium, or a computer-readable medium) comprising, recorded thereon, the computer program for performing one of the methods described herein.
- the data carrier, the digital storage medium or the recorded medium are typically tangible and/or non- transitionary.
- a further embodiment of the inventive method is, therefore, a data stream or a sequence of signals representing the computer program for performing one of the methods described herein.
- the data stream or the sequence of signals may for example be configured to be transferred via a data communication connection, for example via the Internet.
- a further embodiment comprises a processing means, for example a computer, or a programmable logic device, configured to or adapted to perform one of the methods described herein.
- a further embodiment comprises a computer having installed thereon the computer program for performing one of the methods described herein.
- a further embodiment according to the invention comprises an apparatus or a system configured to transfer (for example, electronically or optically) a computer program for performing one of the methods described herein to a receiver.
- the receiver may, for example, be a computer, a mobile device, a memory device or the like.
- the apparatus or system may, for example, comprise a file server for transferring the computer program to the receiver.
- a programmable logic device for example a field programmable gate array
- a field programmable gate array may cooperate with a microprocessor in order to perform one of the methods described herein.
- the methods are preferably performed by any hardware apparatus.
- the apparatus described herein may be implemented using a hardware apparatus, or using a computer, or using a combination of a hardware apparatus and a computer.
- the apparatus described herein, or any components of the apparatus described herein, may be implemented at least partially in hardware and/or in software.
- the methods described herein may be performed using a hardware apparatus, or using a computer, or using a combination of a hardware apparatus and a computer.
- ITU-T Perceptual evaluation of speech quality (PESQ), an objective method for end- to-end speech quality assessment of narrowband telephone networks and speech codecs, International Telecommunications Union (ITU-T) Recommendation P.862, Feb. 2001 .
- PESQ Perceptual evaluation of speech quality
Landscapes
- Engineering & Computer Science (AREA)
- Computational Linguistics (AREA)
- Quality & Reliability (AREA)
- Signal Processing (AREA)
- Health & Medical Sciences (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Human Computer Interaction (AREA)
- Physics & Mathematics (AREA)
- Acoustics & Sound (AREA)
- Multimedia (AREA)
- Circuit For Audible Band Transducer (AREA)
- Cable Transmission Systems, Equalization Of Radio And Reduction Of Echo (AREA)
Abstract
L'invention concerne un processeur de signal destiné à fournir un ou plusieurs signaux audio traités sur la base d'un ou de plusieurs signaux audio d'entrée, ledit processeur de signal étant configuré pour estimer des coefficients d'un modèle de réverbération autorégressif à l'aide des signaux audio d'entrée et des signaux réverbérés à bruit réduit retardés obtenus par réduction de bruit. Le processeur de signal est configuré pour fournir des signaux réverbérés à bruit réduit à l'aide des signaux audio d'entrée et des coefficients estimés du modèle de réverbération autorégressif. Le processeur de signal est configuré pour dériver des signaux de sortie à bruit réduit et à réverbération réduite à l'aide des signaux réverbérés à bruit réduit et des coefficients estimés du modèle de réverbération autorégressif. La présente invention concerne également un procédé et un programme d'ordinateur comprenant une fonctionnalité similaire.A signal processor for providing one or more processed audio signals based on one or more input audio signals, wherein said signal processor is configured to estimate coefficients of an autoregressive reverb model to using input audio signals and delayed noise reduced reverberated signals obtained by noise reduction. The signal processor is configured to provide reduced noise reverberated signals using the input audio signals and estimated coefficients of the autoregressive reverb model. The signal processor is configured to derive reduced noise, reduced reverberation output signals using the reduced noise reverberated signals and estimated coefficients of the autoregressive reverberation model. The present invention also relates to a method and a computer program comprising a similar functionality.
Description
Claims
Applications Claiming Priority (3)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| EP17192396 | 2017-09-21 | ||
| EP18158479.8A EP3460795A1 (en) | 2017-09-21 | 2018-02-23 | Signal processor and method for providing a processed audio signal reducing noise and reverberation |
| PCT/EP2018/075529 WO2019057847A1 (en) | 2017-09-21 | 2018-09-20 | Signal processor and method for providing a processed audio signal reducing noise and reverberation |
Publications (2)
| Publication Number | Publication Date |
|---|---|
| EP3685378A1 true EP3685378A1 (en) | 2020-07-29 |
| EP3685378B1 EP3685378B1 (en) | 2021-10-13 |
Family
ID=60001661
Family Applications (2)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP18158479.8A Withdrawn EP3460795A1 (en) | 2017-09-21 | 2018-02-23 | Signal processor and method for providing a processed audio signal reducing noise and reverberation |
| EP18769221.5A Active EP3685378B1 (en) | 2017-09-21 | 2018-09-20 | Signal processor and method for providing a processed audio signal reducing noise and reverberation |
Family Applications Before (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| EP18158479.8A Withdrawn EP3460795A1 (en) | 2017-09-21 | 2018-02-23 | Signal processor and method for providing a processed audio signal reducing noise and reverberation |
Country Status (7)
| Country | Link |
|---|---|
| US (1) | US11133019B2 (en) |
| EP (2) | EP3460795A1 (en) |
| JP (1) | JP6894580B2 (en) |
| CN (1) | CN111512367B (en) |
| BR (1) | BR112020005809A2 (en) |
| RU (1) | RU2768514C2 (en) |
| WO (1) | WO2019057847A1 (en) |
Families Citing this family (13)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CA3147429A1 (en) | 2019-08-01 | 2021-02-04 | Dolby Laboratories Licensing Corporation | Systems and methods for covariance smoothing |
| CN111933170B (en) * | 2020-07-20 | 2024-03-29 | 歌尔科技有限公司 | Voice signal processing method, device, equipment and storage medium |
| US12597432B2 (en) | 2020-07-30 | 2026-04-07 | Dolby International Ab | Hum noise detection and removal for speech and music recordings |
| CN112017680B (en) * | 2020-08-26 | 2024-07-02 | 西北工业大学 | Dereverberation method and device |
| CN112017682B (en) * | 2020-09-18 | 2023-05-23 | 中科极限元(杭州)智能科技股份有限公司 | Single-channel voice simultaneous noise reduction and reverberation removal system |
| EP3971892A1 (en) * | 2020-09-18 | 2022-03-23 | Fraunhofer-Gesellschaft zur Förderung der angewandten Forschung e.V. | Apparatus and method for combining repeated noisy signals |
| CN113160842B (en) * | 2021-03-06 | 2024-04-09 | 西安电子科技大学 | A speech dereverberation method and system based on MCLP |
| CN113115196B (en) * | 2021-04-22 | 2022-03-29 | 东莞市声强电子有限公司 | Intelligent test method of noise reduction earphone |
| US12456456B2 (en) | 2022-01-20 | 2025-10-28 | Microsoft Technology Licensing, Llc | Data augmentation system and method for multi-microphone systems |
| US12579984B2 (en) * | 2022-01-20 | 2026-03-17 | Microsoft Technology Licensing, Llc. | Data augmentation system and method for multi-microphone systems |
| CN114928659B (en) * | 2022-07-20 | 2022-09-30 | 深圳市子恒通讯设备有限公司 | Exhaust silencing method for multiplex communication |
| CN118430560B (en) * | 2024-05-06 | 2025-10-03 | 厦门亿联网络技术股份有限公司 | A noise elimination method, system, device and medium based on bipolar Kalman |
| US20260024538A1 (en) * | 2024-07-17 | 2026-01-22 | Tencent America LLC | Innovative karaoke sound system with combined neural network feedback suppression and coded vocal restoration |
Family Cites Families (15)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| SE506034C2 (en) * | 1996-02-01 | 1997-11-03 | Ericsson Telefon Ab L M | Method and apparatus for improving parameters representing noise speech |
| JP3986457B2 (en) * | 2003-03-28 | 2007-10-03 | 日本電信電話株式会社 | Input signal estimation method and apparatus, input signal estimation program, and recording medium therefor |
| US8467538B2 (en) | 2008-03-03 | 2013-06-18 | Nippon Telegraph And Telephone Corporation | Dereverberation apparatus, dereverberation method, dereverberation program, and recording medium |
| JP4880036B2 (en) | 2006-05-01 | 2012-02-22 | 日本電信電話株式会社 | Method and apparatus for speech dereverberation based on stochastic model of sound source and room acoustics |
| EP2058804B1 (en) * | 2007-10-31 | 2016-12-14 | Nuance Communications, Inc. | Method for dereverberation of an acoustic signal and system thereof |
| US8848933B2 (en) * | 2008-03-06 | 2014-09-30 | Nippon Telegraph And Telephone Corporation | Signal enhancement device, method thereof, program, and recording medium |
| JP4977100B2 (en) * | 2008-08-11 | 2012-07-18 | 日本電信電話株式会社 | Reverberation removal apparatus, dereverberation removal method, program thereof, and recording medium |
| JP5709760B2 (en) | 2008-12-18 | 2015-04-30 | コーニンクレッカ フィリップス エヌ ヴェ | Audio noise canceling |
| CN101477801B (en) * | 2009-01-22 | 2012-01-04 | 东华大学 | Method for detecting and eliminating pulse noise in digital audio signal |
| DK2463856T3 (en) * | 2010-12-09 | 2014-09-22 | Oticon As | Method of reducing artifacts in algorithms with rapidly varying amplification |
| EP2541542A1 (en) * | 2011-06-27 | 2013-01-02 | Fraunhofer-Gesellschaft zur Förderung der angewandten Forschung e.V. | Apparatus and method for determining a measure for a perceived level of reverberation, audio processor and method for processing a signal |
| JP5897343B2 (en) | 2012-02-17 | 2016-03-30 | 株式会社日立製作所 | Reverberation parameter estimation apparatus and method, dereverberation / echo cancellation parameter estimation apparatus, dereverberation apparatus, dereverberation / echo cancellation apparatus, and dereverberation apparatus online conference system |
| CN102750956B (en) * | 2012-06-18 | 2014-07-16 | 歌尔声学股份有限公司 | Method and device for removing reverberation of single channel voice |
| DK3190587T3 (en) * | 2012-08-24 | 2019-01-21 | Oticon As | Noise estimation for noise reduction and echo suppression in personal communication |
| EP2747451A1 (en) * | 2012-12-21 | 2014-06-25 | Fraunhofer-Gesellschaft zur Förderung der angewandten Forschung e.V. | Filter and method for informed spatial filtering using multiple instantaneous direction-of-arrivial estimates |
-
2018
- 2018-02-23 EP EP18158479.8A patent/EP3460795A1/en not_active Withdrawn
- 2018-09-20 EP EP18769221.5A patent/EP3685378B1/en active Active
- 2018-09-20 WO PCT/EP2018/075529 patent/WO2019057847A1/en not_active Ceased
- 2018-09-20 JP JP2020516618A patent/JP6894580B2/en active Active
- 2018-09-20 CN CN201880073959.4A patent/CN111512367B/en active Active
- 2018-09-20 BR BR112020005809-2A patent/BR112020005809A2/en unknown
- 2018-09-20 RU RU2020113933A patent/RU2768514C2/en active
-
2020
- 2020-03-19 US US16/824,421 patent/US11133019B2/en active Active
Also Published As
| Publication number | Publication date |
|---|---|
| US20200219524A1 (en) | 2020-07-09 |
| CN111512367B (en) | 2023-03-14 |
| EP3460795A1 (en) | 2019-03-27 |
| RU2768514C2 (en) | 2022-03-24 |
| EP3685378B1 (en) | 2021-10-13 |
| RU2020113933A (en) | 2021-10-21 |
| RU2020113933A3 (en) | 2021-10-21 |
| CN111512367A (en) | 2020-08-07 |
| JP2020537172A (en) | 2020-12-17 |
| JP6894580B2 (en) | 2021-06-30 |
| BR112020005809A2 (en) | 2020-09-24 |
| US11133019B2 (en) | 2021-09-28 |
| WO2019057847A1 (en) | 2019-03-28 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| EP3685378A1 (en) | Signal processor and method for providing a processed audio signal reducing noise and reverberation | |
| US11315587B2 (en) | Signal processor for signal enhancement and associated methods | |
| JP7094340B2 (en) | A method for enhancing telephone audio signals based on convolutional neural networks | |
| CN108172231B (en) | A Kalman Filter-Based Reverberation Method and System | |
| Braun et al. | Linear prediction-based online dereverberation and noise reduction using alternating Kalman filters | |
| CA3124017C (en) | Apparatus and method for source separation using an estimation and control of sound quality | |
| CN108141656B (en) | Method and apparatus for digital signal processing of microphones | |
| US8521530B1 (en) | System and method for enhancing a monaural audio signal | |
| JP5645419B2 (en) | Reverberation removal device | |
| CN120932662A (en) | System and method for enhancing degraded audio signals | |
| KR102076760B1 (en) | Method for cancellating nonlinear acoustic echo based on kalman filtering using microphone array | |
| US20200286501A1 (en) | Apparatus and a method for signal enhancement | |
| Zhou et al. | Speech dereverberation with a reverberation time shortening target | |
| GB2577905A (en) | Processing audio signals | |
| Parchami et al. | Speech dereverberation using linear prediction with estimation of early speech spectral variance | |
| JP3673727B2 (en) | Reverberation elimination method, apparatus thereof, program thereof, and recording medium thereof | |
| Lohmann et al. | Dereverberation in acoustic sensor networks using weighted prediction error with microphone-dependent prediction delays | |
| Rahmani et al. | An iterative noise cross-PSD estimation for two-microphone speech enhancement | |
| KR20200095370A (en) | Detection of fricatives in speech signals | |
| Li et al. | Joint noise reduction and listening enhancement for full-end speech enhancement | |
| KR102056398B1 (en) | Real-time speech derverberation method and apparatus using multi-channel linear prediction with estimation of early speech psd for distant speech recognition | |
| Li et al. | Adaptive dereverberation using multi-channel linear prediction with deficient length filter | |
| Lemercier et al. | Neural Network-augmented Kalman Filtering for Robust Online Speech Dereverberation in Noisy Reverberant Environments | |
| RU2782364C1 (en) | Apparatus and method for isolating sources using sound quality assessment and control | |
| Nathwani et al. | Multi channel reverberant speech enhancement using LP residual cepstrum |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: UNKNOWN |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE INTERNATIONAL PUBLICATION HAS BEEN MADE |
|
| PUAI | Public reference made under article 153(3) epc to a published international application that has entered the european phase |
Free format text: ORIGINAL CODE: 0009012 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: REQUEST FOR EXAMINATION WAS MADE |
|
| 17P | Request for examination filed |
Effective date: 20200320 |
|
| AK | Designated contracting states |
Kind code of ref document: A1 Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC MK MT NL NO PL PT RO RS SE SI SK SM TR |
|
| AX | Request for extension of the european patent |
Extension state: BA ME |
|
| RIN1 | Information on inventor provided before grant (corrected) |
Inventor name: HABETS, EMANUEL Inventor name: BRAUN, SEBASTIAN |
|
| DAV | Request for validation of the european patent (deleted) | ||
| DAX | Request for extension of the european patent (deleted) | ||
| GRAP | Despatch of communication of intention to grant a patent |
Free format text: ORIGINAL CODE: EPIDOSNIGR1 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: GRANT OF PATENT IS INTENDED |
|
| INTG | Intention to grant announced |
Effective date: 20210422 |
|
| RAP3 | Party data changed (applicant data changed or rights of an application transferred) |
Owner name: FRAUNHOFER-GESELLSCHAFT ZUR FOERDERUNG DER ANGEWANDTEN FORSCHUNG E.V. |
|
| GRAS | Grant fee paid |
Free format text: ORIGINAL CODE: EPIDOSNIGR3 |
|
| GRAA | (expected) grant |
Free format text: ORIGINAL CODE: 0009210 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: THE PATENT HAS BEEN GRANTED |
|
| AK | Designated contracting states |
Kind code of ref document: B1 Designated state(s): AL AT BE BG CH CY CZ DE DK EE ES FI FR GB GR HR HU IE IS IT LI LT LU LV MC MK MT NL NO PL PT RO RS SE SI SK SM TR |
|
| REG | Reference to a national code |
Ref country code: GB Ref legal event code: FG4D |
|
| REG | Reference to a national code |
Ref country code: CH Ref legal event code: EP |
|
| REG | Reference to a national code |
Ref country code: DE Ref legal event code: R096 Ref document number: 602018025052 Country of ref document: DE |
|
| REG | Reference to a national code |
Ref country code: IE Ref legal event code: FG4D |
|
| REG | Reference to a national code |
Ref country code: AT Ref legal event code: REF Ref document number: 1438769 Country of ref document: AT Kind code of ref document: T Effective date: 20211115 |
|
| REG | Reference to a national code |
Ref country code: LT Ref legal event code: MG9D |
|
| REG | Reference to a national code |
Ref country code: NL Ref legal event code: MP Effective date: 20211013 |
|
| REG | Reference to a national code |
Ref country code: AT Ref legal event code: MK05 Ref document number: 1438769 Country of ref document: AT Kind code of ref document: T Effective date: 20211013 |
|
| RAP4 | Party data changed (patent owner data changed or rights of a patent transferred) |
Owner name: FRAUNHOFER-GESELLSCHAFT ZUR FOERDERUNG DER ANGEWANDTEN FORSCHUNG E.V. |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: RS Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: LT Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: FI Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: BG Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20220113 Ref country code: AT Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: IS Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20220213 Ref country code: SE Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: PT Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20220214 Ref country code: PL Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: NO Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20220113 Ref country code: NL Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: LV Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: HR Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: GR Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20220114 Ref country code: ES Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| REG | Reference to a national code |
Ref country code: DE Ref legal event code: R097 Ref document number: 602018025052 Country of ref document: DE |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: SM Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: SK Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: RO Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: EE Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: DK Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: CZ Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| PLBE | No opposition filed within time limit |
Free format text: ORIGINAL CODE: 0009261 |
|
| STAA | Information on the status of an ep patent application or granted ep patent |
Free format text: STATUS: NO OPPOSITION FILED WITHIN TIME LIMIT |
|
| 26N | No opposition filed |
Effective date: 20220714 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: AL Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: SI Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: MC Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| REG | Reference to a national code |
Ref country code: CH Ref legal event code: PL |
|
| REG | Reference to a national code |
Ref country code: BE Ref legal event code: MM Effective date: 20220930 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: IT Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| P01 | Opt-out of the competence of the unified patent court (upc) registered |
Effective date: 20230517 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: LU Free format text: LAPSE BECAUSE OF NON-PAYMENT OF DUE FEES Effective date: 20220920 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: LI Free format text: LAPSE BECAUSE OF NON-PAYMENT OF DUE FEES Effective date: 20220930 Ref country code: IE Free format text: LAPSE BECAUSE OF NON-PAYMENT OF DUE FEES Effective date: 20220920 Ref country code: CH Free format text: LAPSE BECAUSE OF NON-PAYMENT OF DUE FEES Effective date: 20220930 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: BE Free format text: LAPSE BECAUSE OF NON-PAYMENT OF DUE FEES Effective date: 20220930 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: CY Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: MK Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 Ref country code: HU Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT; INVALID AB INITIO Effective date: 20180920 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: TR Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| PG25 | Lapsed in a contracting state [announced via postgrant information from national office to epo] |
Ref country code: MT Free format text: LAPSE BECAUSE OF FAILURE TO SUBMIT A TRANSLATION OF THE DESCRIPTION OR TO PAY THE FEE WITHIN THE PRESCRIBED TIME-LIMIT Effective date: 20211013 |
|
| PGFP | Annual fee paid to national office [announced via postgrant information from national office to epo] |
Ref country code: DE Payment date: 20240919 Year of fee payment: 7 |
|
| PGFP | Annual fee paid to national office [announced via postgrant information from national office to epo] |
Ref country code: GB Payment date: 20240923 Year of fee payment: 7 |
|
| PGFP | Annual fee paid to national office [announced via postgrant information from national office to epo] |
Ref country code: FR Payment date: 20240924 Year of fee payment: 7 |