PL3594948T3 - Klasyfikator sygnału audio - Google Patents

Klasyfikator sygnału audio

Info

Publication number
PL3594948T3
PL3594948T3 PL19195287T PL19195287T PL3594948T3 PL 3594948 T3 PL3594948 T3 PL 3594948T3 PL 19195287 T PL19195287 T PL 19195287T PL 19195287 T PL19195287 T PL 19195287T PL 3594948 T3 PL3594948 T3 PL 3594948T3
Authority
PL
Poland
Prior art keywords
audio signal
signal classifier
classifier
audio
signal
Prior art date
Application number
PL19195287T
Other languages
English (en)
Inventor
Erik Norvell
Volodya Grancharov
Original Assignee
Telefonaktiebolaget Lm Ericsson (Publ)
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Telefonaktiebolaget Lm Ericsson (Publ) filed Critical Telefonaktiebolaget Lm Ericsson (Publ)
Publication of PL3594948T3 publication Critical patent/PL3594948T3/pl

Links

Classifications

    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/48—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 specially adapted for particular use
    • G10L25/51—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 specially adapted for particular use for comparison or discrimination
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
    • G10L19/06—Determination or coding of the spectral characteristics, e.g. of the short-term prediction coefficients
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
    • G10L19/16—Vocoder architecture
    • G10L19/167—Audio streaming, i.e. formatting and decoding of an encoded audio signal representation into a data stream for transmission or storage purposes
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
    • G10L19/16—Vocoder architecture
    • G10L19/18—Vocoders using multiple modes
    • G10L19/22—Mode decision, i.e. based on audio signal content versus external parameters
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/03—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters
    • G10L25/18—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 characterised by the type of extracted parameters the extracted parameters being spectral information of each sub-band
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
    • G10L25/78—Detection of presence or absence of voice signals
    • G10L25/81—Detection of presence or absence of voice signals for discriminating voice from music
    • G—PHYSICS
    • G10—MUSICAL INSTRUMENTS; ACOUSTICS
    • G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
    • G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
    • G10L19/16—Vocoder architecture
    • G10L19/18—Vocoders using multiple modes
    • G10L19/20—Vocoders using multiple modes using sound class specific coding, hybrid encoders or object based coding

Landscapes

  • Engineering & Computer Science (AREA)
  • Physics & Mathematics (AREA)
  • Computational Linguistics (AREA)
  • Signal Processing (AREA)
  • Health & Medical Sciences (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Human Computer Interaction (AREA)
  • Acoustics & Sound (AREA)
  • Multimedia (AREA)
  • Spectroscopy & Molecular Physics (AREA)
  • Compression, Expansion, Code Conversion, And Decoders (AREA)
PL19195287T 2014-05-08 2015-05-07 Klasyfikator sygnału audio PL3594948T3 (pl)

Applications Claiming Priority (5)

Application Number Priority Date Filing Date Title
US201461990354P 2014-05-08 2014-05-08
PCT/SE2015/050503 WO2015171061A1 (en) 2014-05-08 2015-05-07 Audio signal discriminator and coder
EP18172361.0A EP3379535B1 (en) 2014-05-08 2015-05-07 Audio signal classifier
EP15724098.7A EP3140831B1 (en) 2014-05-08 2015-05-07 Audio signal discriminator and coder
EP19195287.8A EP3594948B1 (en) 2014-05-08 2015-05-07 Audio signal classifier

Publications (1)

Publication Number Publication Date
PL3594948T3 true PL3594948T3 (pl) 2021-08-30

Family

ID=53200274

Family Applications (2)

Application Number Title Priority Date Filing Date
PL15724098T PL3140831T3 (pl) 2014-05-08 2015-05-07 Dyskryminator i koder sygnału audio
PL19195287T PL3594948T3 (pl) 2014-05-08 2015-05-07 Klasyfikator sygnału audio

Family Applications Before (1)

Application Number Title Priority Date Filing Date
PL15724098T PL3140831T3 (pl) 2014-05-08 2015-05-07 Dyskryminator i koder sygnału audio

Country Status (11)

Country Link
US (3) US9620138B2 (pl)
EP (3) EP3140831B1 (pl)
CN (3) CN110619891B (pl)
BR (1) BR112016025850B1 (pl)
DK (2) DK3140831T3 (pl)
ES (3) ES2763280T3 (pl)
HU (1) HUE046477T2 (pl)
MX (2) MX356883B (pl)
MY (1) MY182165A (pl)
PL (2) PL3140831T3 (pl)
WO (1) WO2015171061A1 (pl)

Families Citing this family (7)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
RU2750644C2 (ru) 2013-10-18 2021-06-30 Телефонактиеболагет Л М Эрикссон (Пабл) Кодирование и декодирование положений спектральных пиков
EP3140831B1 (en) * 2014-05-08 2018-07-11 Telefonaktiebolaget LM Ericsson (publ) Audio signal discriminator and coder
PL3163571T3 (pl) * 2014-07-28 2020-05-18 Nippon Telegraph And Telephone Corporation Kodowanie sygnału dźwiękowego
CN110211580B (zh) * 2019-05-15 2021-07-16 海尔优家智能科技(北京)有限公司 多智能设备应答方法、装置、系统及存储介质
CA3184152A1 (en) * 2020-06-30 2022-01-06 Rivarol VERGIN Cumulative average spectral entropy analysis for tone and speech classification
CN113890492B (zh) * 2021-10-09 2025-07-18 深圳市创成微电子有限公司 音频功率放大器的供电电压控制方法、控制器和音频设备
US20250201255A1 (en) * 2023-12-13 2025-06-19 Qualcomm Incorporated Content-based switchable audio codec

Family Cites Families (26)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
ATE384358T1 (de) * 1998-05-27 2008-02-15 Microsoft Corp Verfahren und vorrichtung zur maskierung des quantisierungsrauschens von audiosignalen
US6226608B1 (en) * 1999-01-28 2001-05-01 Dolby Laboratories Licensing Corporation Data framing for adaptive-block-length coding system
US6959274B1 (en) * 1999-09-22 2005-10-25 Mindspeed Technologies, Inc. Fixed rate speech compression system and method
US6785645B2 (en) * 2001-11-29 2004-08-31 Microsoft Corporation Real-time speech and music classifier
KR100762596B1 (ko) * 2006-04-05 2007-10-01 삼성전자주식회사 음성 신호 전처리 시스템 및 음성 신호 특징 정보 추출방법
US20070282601A1 (en) * 2006-06-02 2007-12-06 Texas Instruments Inc. Packet loss concealment for a conjugate structure algebraic code excited linear prediction decoder
CN101145345B (zh) * 2006-09-13 2011-02-09 华为技术有限公司 音频分类方法
ES2533358T3 (es) * 2007-06-22 2015-04-09 Voiceage Corporation Procedimiento y dispositivo para estimar la tonalidad de una señal de sonido
CN101399039B (zh) * 2007-09-30 2011-05-11 华为技术有限公司 一种确定非噪声音频信号类别的方法及装置
KR101599875B1 (ko) * 2008-04-17 2016-03-14 삼성전자주식회사 멀티미디어의 컨텐트 특성에 기반한 멀티미디어 부호화 방법 및 장치, 멀티미디어의 컨텐트 특성에 기반한 멀티미디어 복호화 방법 및 장치
EP2346029B1 (en) * 2008-07-11 2013-06-05 Fraunhofer-Gesellschaft zur Förderung der angewandten Forschung e.V. Audio encoder, method for encoding an audio signal and corresponding computer program
EP2210944A1 (en) 2009-01-22 2010-07-28 ATG:biosynthetics GmbH Methods for generation of RNA and (poly)peptide libraries and their use
CN102044246B (zh) * 2009-10-15 2012-05-23 华为技术有限公司 一种音频信号检测方法和装置
KR101754970B1 (ko) * 2010-01-12 2017-07-06 삼성전자주식회사 무선 통신 시스템의 채널 상태 측정 기준신호 처리 장치 및 방법
US9652999B2 (en) * 2010-04-29 2017-05-16 Educational Testing Service Computer-implemented systems and methods for estimating word accuracy for automatic speech recognition
US8977542B2 (en) * 2010-07-16 2015-03-10 Telefonaktiebolaget L M Ericsson (Publ) Audio encoder and decoder and methods for encoding and decoding an audio signal
RU2010152225A (ru) * 2010-12-20 2012-06-27 ЭлЭсАй Корпорейшн (US) Обнаружение музыки с использованием анализа спектральных пиков
CN102982804B (zh) * 2011-09-02 2017-05-03 杜比实验室特许公司 音频分类方法和系统
CN102522082B (zh) * 2011-12-27 2013-07-10 重庆大学 一种公共场所异常声音的识别与定位方法
US9111531B2 (en) * 2012-01-13 2015-08-18 Qualcomm Incorporated Multiple coding mode signal classification
US20130282372A1 (en) * 2012-04-23 2013-10-24 Qualcomm Incorporated Systems and methods for audio signal processing
AU2013283568B2 (en) * 2012-06-28 2016-05-12 Fraunhofer-Gesellschaft Zur Foerderung Der Angewandten Forschung E.V. Linear prediction based audio coding using improved probability distribution estimation
US9401153B2 (en) * 2012-10-15 2016-07-26 Digimarc Corporation Multi-mode audio recognition and auxiliary data encoding and decoding
EP3140831B1 (en) * 2014-05-08 2018-07-11 Telefonaktiebolaget LM Ericsson (publ) Audio signal discriminator and coder
WO2015168925A1 (en) 2014-05-09 2015-11-12 Qualcomm Incorporated Restricted aperiodic csi measurement reporting in enhanced interference management and traffic adaptation
TWI602172B (zh) * 2014-08-27 2017-10-11 弗勞恩霍夫爾協會 使用參數以加強隱蔽之用於編碼及解碼音訊內容的編碼器、解碼器及方法

Also Published As

Publication number Publication date
US20190198032A1 (en) 2019-06-27
CN110619892B (zh) 2023-04-11
EP3594948A1 (en) 2020-01-15
MX2016014534A (es) 2017-02-20
EP3379535B1 (en) 2019-09-18
ES2690577T3 (es) 2018-11-21
ES2763280T3 (es) 2020-05-27
MY182165A (en) 2021-01-18
US20170178660A1 (en) 2017-06-22
BR112016025850A2 (pl) 2017-08-15
CN110619891B (zh) 2023-01-17
MX2018007257A (es) 2022-08-25
CN106463141B (zh) 2019-11-01
DK3379535T3 (da) 2019-12-16
US10242687B2 (en) 2019-03-26
CN110619891A (zh) 2019-12-27
PL3140831T3 (pl) 2018-12-31
BR112016025850B1 (pt) 2022-08-16
EP3379535A1 (en) 2018-09-26
US20160086615A1 (en) 2016-03-24
EP3140831A1 (en) 2017-03-15
EP3140831B1 (en) 2018-07-11
US10984812B2 (en) 2021-04-20
CN106463141A (zh) 2017-02-22
MX356883B (es) 2018-06-19
EP3594948B1 (en) 2021-03-03
WO2015171061A1 (en) 2015-11-12
HUE046477T2 (hu) 2020-03-30
CN110619892A (zh) 2019-12-27
US9620138B2 (en) 2017-04-11
ES2874757T3 (es) 2021-11-05
DK3140831T3 (en) 2018-10-15

Similar Documents

Publication Publication Date Title
GB201406574D0 (en) Audio Signal Processing
GB2556015B (en) Audio Signals
GB201518004D0 (en) Audio signal processing
GB201405123D0 (en) Audio signal payload
PT3097397T (pt) Sistema de medição de carris
GB201401689D0 (en) Audio signal processing
GB201401626D0 (en) Audio signal analysis
GB201409766D0 (en) Signal processing methods
GB2533373B (en) Video-based sound source separation
GB201503467D0 (en) Analysing audio data
GB2549805B (en) Audio signals
TWM489443U (en) Headphone
HUE046477T2 (hu) Hangjel-osztályozó
GB201717281D0 (en) Audio Signal
GB2538450B (en) Signal system
TWM490712U (en) Improved speaker structure
EP3140998A4 (en) Speaker
PL3790007T3 (pl) Kodowanie audio
EP3095117A4 (en) Multi-channel audio signal classifier
GB2524901B (en) Loudspeaker
EP3127621A4 (en) Classifier
TWI561093B (en) Speaker structure
GB2522055B (en) Loudspeaker system
GB2551605B (en) Audio signal processor
GB201407263D0 (en) Loudspeaker