IN2014MN01588A - - Google Patents
Info
- Publication number
- IN2014MN01588A IN2014MN01588A IN1588MUN2014A IN2014MN01588A IN 2014MN01588 A IN2014MN01588 A IN 2014MN01588A IN 1588MUN2014 A IN1588MUN2014 A IN 1588MUN2014A IN 2014MN01588 A IN2014MN01588 A IN 2014MN01588A
- Authority
- IN
- India
- Prior art keywords
- classification
- speech
- frame
- music
- classified
- Prior art date
Links
- 230000000694 effects Effects 0.000 abstract 1
- 230000007774 longterm Effects 0.000 abstract 1
Classifications
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L17/00—Speaker identification or verification techniques
- G10L17/02—Preprocessing operations, e.g. segment selection; Pattern representation or modelling, e.g. based on linear discriminant analysis [LDA] or principal components; Feature selection or extraction
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
- G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
- G10L19/16—Vocoder architecture
- G10L19/18—Vocoders using multiple modes
- G10L19/20—Vocoders using multiple modes using sound class specific coding, hybrid encoders or object based coding
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
- G10L19/02—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using spectral analysis, e.g. transform vocoders or subband vocoders
- G10L19/0212—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using spectral analysis, e.g. transform vocoders or subband vocoders using orthogonal transformation
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
- G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
- G10L19/08—Determination or coding of the excitation function; Determination or coding of the long-term prediction parameters
- G10L19/12—Determination or coding of the excitation function; Determination or coding of the long-term prediction parameters the excitation function being a code excitation, e.g. in code excited linear prediction [CELP] vocoders
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
- G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
- G10L19/16—Vocoder architecture
- G10L19/173—Transcoding, i.e. converting between two coded representations avoiding cascaded coding-decoding
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
- G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
- G10L19/16—Vocoder architecture
- G10L19/18—Vocoders using multiple modes
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
- G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
- G10L19/16—Vocoder architecture
- G10L19/18—Vocoders using multiple modes
- G10L19/22—Mode decision, i.e. based on audio signal content versus external parameters
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L19/00—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis
- G10L19/04—Speech or audio signals analysis-synthesis techniques for redundancy reduction, e.g. in vocoders; Coding or decoding of speech or audio signals, using source filter models or psychoacoustic analysis using predictive techniques
- G10L19/16—Vocoder architecture
- G10L19/18—Vocoders using multiple modes
- G10L19/24—Variable rate codecs, e.g. for generating different qualities using a scalable representation such as hierarchical encoding or layered encoding
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
- G10L25/78—Detection of presence or absence of voice signals
- G10L25/81—Detection of presence or absence of voice signals for discriminating voice from music
Landscapes
- Engineering & Computer Science (AREA)
- Physics & Mathematics (AREA)
- Health & Medical Sciences (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Human Computer Interaction (AREA)
- Acoustics & Sound (AREA)
- Multimedia (AREA)
- Signal Processing (AREA)
- Computational Linguistics (AREA)
- Computer Vision & Pattern Recognition (AREA)
- Spectroscopy & Molecular Physics (AREA)
- Quality & Reliability (AREA)
- Compression, Expansion, Code Conversion, And Decoders (AREA)
Applications Claiming Priority (3)
Application Number | Priority Date | Filing Date | Title |
---|---|---|---|
US201261586374P | 2012-01-13 | 2012-01-13 | |
US13/722,669 US9111531B2 (en) | 2012-01-13 | 2012-12-20 | Multiple coding mode signal classification |
PCT/US2012/071217 WO2013106192A1 (en) | 2012-01-13 | 2012-12-21 | Multiple coding mode signal classification |
Publications (1)
Publication Number | Publication Date |
---|---|
IN2014MN01588A true IN2014MN01588A (enrdf_load_stackoverflow) | 2015-05-08 |
Family
ID=48780608
Family Applications (1)
Application Number | Title | Priority Date | Filing Date |
---|---|---|---|
IN1588MUN2014 IN2014MN01588A (enrdf_load_stackoverflow) | 2012-01-13 | 2012-12-21 |
Country Status (12)
Families Citing this family (18)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
US9589570B2 (en) * | 2012-09-18 | 2017-03-07 | Huawei Technologies Co., Ltd. | Audio classification based on perceptual quality for low or medium bit rates |
EP3933836B1 (en) * | 2012-11-13 | 2024-07-31 | Samsung Electronics Co., Ltd. | Method and apparatus for determining encoding mode, method and apparatus for encoding audio signals, and method and apparatus for decoding audio signals |
CN106409310B (zh) | 2013-08-06 | 2019-11-19 | 华为技术有限公司 | 一种音频信号分类方法和装置 |
CN104424956B9 (zh) | 2013-08-30 | 2022-11-25 | 中兴通讯股份有限公司 | 激活音检测方法和装置 |
US10090004B2 (en) * | 2014-02-24 | 2018-10-02 | Samsung Electronics Co., Ltd. | Signal classifying method and device, and audio encoding method and device using same |
CN110619892B (zh) * | 2014-05-08 | 2023-04-11 | 瑞典爱立信有限公司 | 音频信号区分器和编码器 |
CN107424622B (zh) | 2014-06-24 | 2020-12-25 | 华为技术有限公司 | 音频编码方法和装置 |
CN106448688B (zh) * | 2014-07-28 | 2019-11-05 | 华为技术有限公司 | 音频编码方法及相关装置 |
US9886963B2 (en) * | 2015-04-05 | 2018-02-06 | Qualcomm Incorporated | Encoder selection |
CN104867492B (zh) * | 2015-05-07 | 2019-09-03 | 科大讯飞股份有限公司 | 智能交互系统及方法 |
KR102398124B1 (ko) * | 2015-08-11 | 2022-05-17 | 삼성전자주식회사 | 음향 데이터의 적응적 처리 |
US10186276B2 (en) * | 2015-09-25 | 2019-01-22 | Qualcomm Incorporated | Adaptive noise suppression for super wideband music |
US10902043B2 (en) | 2016-01-03 | 2021-01-26 | Gracenote, Inc. | Responding to remote media classification queries using classifier models and context parameters |
WO2017117234A1 (en) * | 2016-01-03 | 2017-07-06 | Gracenote, Inc. | Responding to remote media classification queries using classifier models and context parameters |
JP6996185B2 (ja) * | 2017-09-15 | 2022-01-17 | 富士通株式会社 | 発話区間検出装置、発話区間検出方法及び発話区間検出用コンピュータプログラム |
CN113748461A (zh) | 2019-04-18 | 2021-12-03 | 杜比实验室特许公司 | 对话检测器 |
EP4421804A4 (en) * | 2021-10-21 | 2024-10-30 | Beijing Xiaomi Mobile Software Co., Ltd. | METHOD AND DEVICE FOR SIGNAL CODING AND DECODING AS WELL AS CODING DEVICE, DECODING DEVICE AND STORAGE MEDIUM |
CN116149499B (zh) * | 2023-04-18 | 2023-08-11 | 深圳雷柏科技股份有限公司 | 用于鼠标的多模式切换控制电路及切换控制方法 |
Family Cites Families (39)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
ES2348319T3 (es) * | 1991-06-11 | 2010-12-02 | Qualcomm Incorporated | Vocodificador de velocidad variable. |
US5778335A (en) | 1996-02-26 | 1998-07-07 | The Regents Of The University Of California | Method and apparatus for efficient multiband celp wideband speech and music coding and decoding |
US6493665B1 (en) * | 1998-08-24 | 2002-12-10 | Conexant Systems, Inc. | Speech classification and parameter weighting used in codebook search |
US7072832B1 (en) * | 1998-08-24 | 2006-07-04 | Mindspeed Technologies, Inc. | System for speech encoding having an adaptive encoding arrangement |
US7272556B1 (en) * | 1998-09-23 | 2007-09-18 | Lucent Technologies Inc. | Scalable and embedded codec for speech and audio signals |
US6691084B2 (en) * | 1998-12-21 | 2004-02-10 | Qualcomm Incorporated | Multiple mode variable rate speech coding |
JP2000267699A (ja) * | 1999-03-19 | 2000-09-29 | Nippon Telegr & Teleph Corp <Ntt> | 音響信号符号化方法および装置、そのプログラム記録媒体、および音響信号復号装置 |
AU6725500A (en) * | 1999-08-23 | 2001-03-19 | Matsushita Electric Industrial Co., Ltd. | Voice encoder and voice encoding method |
US6604070B1 (en) * | 1999-09-22 | 2003-08-05 | Conexant Systems, Inc. | System of encoding and decoding speech signals |
US6625226B1 (en) * | 1999-12-03 | 2003-09-23 | Allen Gersho | Variable bit rate coder, and associated method, for a communication station operable in a communication system |
US6697776B1 (en) * | 2000-07-31 | 2004-02-24 | Mindspeed Technologies, Inc. | Dynamic signal detector system and method |
US6694293B2 (en) | 2001-02-13 | 2004-02-17 | Mindspeed Technologies, Inc. | Speech coding system with a music classifier |
US6785645B2 (en) | 2001-11-29 | 2004-08-31 | Microsoft Corporation | Real-time speech and music classifier |
US6829579B2 (en) * | 2002-01-08 | 2004-12-07 | Dilithium Networks, Inc. | Transcoding method and system between CELP-based speech codes |
US7657427B2 (en) * | 2002-10-11 | 2010-02-02 | Nokia Corporation | Methods and devices for source controlled variable bit-rate wideband speech coding |
US7363218B2 (en) * | 2002-10-25 | 2008-04-22 | Dilithium Networks Pty. Ltd. | Method and apparatus for fast CELP parameter mapping |
FI118834B (fi) * | 2004-02-23 | 2008-03-31 | Nokia Corp | Audiosignaalien luokittelu |
BRPI0418838A (pt) * | 2004-05-17 | 2007-11-13 | Nokia Corp | método para suportar uma codificação de um sinal de áudio, módulo para suportar uma codificação de um sinal de áudio, dispositivo eletrÈnico, sistema de codificação de áudio, e, produto de programa de software |
US8010350B2 (en) | 2006-08-03 | 2011-08-30 | Broadcom Corporation | Decimated bisectional pitch refinement |
CN1920947B (zh) * | 2006-09-15 | 2011-05-11 | 清华大学 | 用于低比特率音频编码的语音/音乐检测器 |
CN101197130B (zh) * | 2006-12-07 | 2011-05-18 | 华为技术有限公司 | 声音活动检测方法和声音活动检测器 |
KR100964402B1 (ko) * | 2006-12-14 | 2010-06-17 | 삼성전자주식회사 | 오디오 신호의 부호화 모드 결정 방법 및 장치와 이를 이용한 오디오 신호의 부호화/복호화 방법 및 장치 |
KR100883656B1 (ko) | 2006-12-28 | 2009-02-18 | 삼성전자주식회사 | 오디오 신호의 분류 방법 및 장치와 이를 이용한 오디오신호의 부호화/복호화 방법 및 장치 |
CN101226744B (zh) * | 2007-01-19 | 2011-04-13 | 华为技术有限公司 | 语音解码器中实现语音解码的方法及装置 |
KR100925256B1 (ko) * | 2007-05-03 | 2009-11-05 | 인하대학교 산학협력단 | 음성 및 음악을 실시간으로 분류하는 방법 |
CN101393741A (zh) * | 2007-09-19 | 2009-03-25 | 中兴通讯股份有限公司 | 一种宽带音频编解码器中的音频信号分类装置及分类方法 |
CN101399039B (zh) * | 2007-09-30 | 2011-05-11 | 华为技术有限公司 | 一种确定非噪声音频信号类别的方法及装置 |
CN101221766B (zh) * | 2008-01-23 | 2011-01-05 | 清华大学 | 音频编码器切换的方法 |
CN101236742B (zh) * | 2008-03-03 | 2011-08-10 | 中兴通讯股份有限公司 | 音乐/非音乐的实时检测方法和装置 |
MX2010009571A (es) * | 2008-03-03 | 2011-05-30 | Lg Electronics Inc | Metodo y aparato para el procesamiento de señales de audio. |
US8768690B2 (en) * | 2008-06-20 | 2014-07-01 | Qualcomm Incorporated | Coding scheme selection for low-bit-rate applications |
WO2010003521A1 (en) * | 2008-07-11 | 2010-01-14 | Fraunhofer-Gesellschaft zur Förderung der angewandten Forschung e.V. | Method and discriminator for classifying different segments of a signal |
EP2144230A1 (en) * | 2008-07-11 | 2010-01-13 | Fraunhofer-Gesellschaft zur Förderung der angewandten Forschung e.V. | Low bitrate audio encoding/decoding scheme having cascaded switches |
KR101261677B1 (ko) * | 2008-07-14 | 2013-05-06 | 광운대학교 산학협력단 | 음성/음악 통합 신호의 부호화/복호화 장치 |
CN101751920A (zh) * | 2008-12-19 | 2010-06-23 | 数维科技(北京)有限公司 | 基于再次分类的音频分类装置及其实现方法 |
CN101814289A (zh) * | 2009-02-23 | 2010-08-25 | 数维科技(北京)有限公司 | 低码率dra数字音频多声道编码方法及其系统 |
JP5519230B2 (ja) * | 2009-09-30 | 2014-06-11 | パナソニック株式会社 | オーディオエンコーダ及び音信号処理システム |
CN102237085B (zh) * | 2010-04-26 | 2013-08-14 | 华为技术有限公司 | 音频信号的分类方法及装置 |
WO2012109734A1 (en) | 2011-02-15 | 2012-08-23 | Voiceage Corporation | Device and method for quantizing the gains of the adaptive and fixed contributions of the excitation in a celp codec |
-
2012
- 2012-12-20 US US13/722,669 patent/US9111531B2/en active Active
- 2012-12-21 SI SI201230593A patent/SI2803068T1/sl unknown
- 2012-12-21 DK DK12810018.7T patent/DK2803068T3/en active
- 2012-12-21 WO PCT/US2012/071217 patent/WO2013106192A1/en active Application Filing
- 2012-12-21 EP EP12810018.7A patent/EP2803068B1/en active Active
- 2012-12-21 CN CN201280066779.6A patent/CN104040626B/zh active Active
- 2012-12-21 ES ES12810018.7T patent/ES2576232T3/es active Active
- 2012-12-21 BR BR112014017001-0A patent/BR112014017001B1/pt active IP Right Grant
- 2012-12-21 JP JP2014552206A patent/JP5964455B2/ja active Active
- 2012-12-21 IN IN1588MUN2014 patent/IN2014MN01588A/en unknown
- 2012-12-21 HU HUE12810018A patent/HUE027037T2/en unknown
- 2012-12-21 KR KR1020147022400A patent/KR20140116487A/ko not_active Ceased
- 2012-12-21 KR KR1020177000172A patent/KR20170005514A/ko not_active Ceased
Also Published As
Publication number | Publication date |
---|---|
BR112014017001A2 (pt) | 2017-06-13 |
US20130185063A1 (en) | 2013-07-18 |
CN104040626B (zh) | 2017-08-11 |
KR20170005514A (ko) | 2017-01-13 |
BR112014017001B1 (pt) | 2020-12-22 |
CN104040626A (zh) | 2014-09-10 |
JP5964455B2 (ja) | 2016-08-03 |
BR112014017001A8 (pt) | 2017-07-04 |
KR20140116487A (ko) | 2014-10-02 |
JP2015507222A (ja) | 2015-03-05 |
WO2013106192A1 (en) | 2013-07-18 |
US9111531B2 (en) | 2015-08-18 |
SI2803068T1 (sl) | 2016-07-29 |
ES2576232T3 (es) | 2016-07-06 |
DK2803068T3 (en) | 2016-05-23 |
HUE027037T2 (en) | 2016-08-29 |
EP2803068B1 (en) | 2016-04-13 |
EP2803068A1 (en) | 2014-11-19 |
Similar Documents
Publication | Publication Date | Title |
---|---|---|
IN2014MN01588A (enrdf_load_stackoverflow) | ||
PH12015501587A1 (en) | Signaling audio rendering information in a bitstream | |
WO2014168934A3 (en) | Systems and methods for generating a digital output signal in a digital microphone system | |
MX2023001960A (es) | Decodificador de audio multicanal, codificador de audio multicanal, metodos y programa de computadora usando un ajuste en base a se?ales residuales de una contribucion de una se?al decorrelacionada. | |
MX351359B (es) | Codificador, decodificador y métodos para la transformación de amplicación por acercamiento dependiente de señales en la codificación espacial de objetos de audio. | |
PH12015501575A1 (en) | Device and method for reducing quantization noise in a time-domain decoder | |
EP4379714A3 (en) | Loudness adjustment for downmixed audio content | |
EP4465662A3 (en) | Compression of decomposed representations of a sound field | |
MX2009007412A (es) | Decodificador de audio. | |
WO2010003109A3 (en) | Speech recognition with parallel recognition tasks | |
EP2846229A3 (en) | Systems and methods for generating haptic effects associated with audio signals | |
MX2016000908A (es) | Aparato y metodo para la codificacion de metadatos de objetos con bajo retardo. | |
IN2015MN01766A (enrdf_load_stackoverflow) | ||
UA116371C2 (uk) | Системи та способи виконання фільтрації для визначення посилення | |
MY187728A (en) | Method and system for encoding audio data with adaptive low frequency compensation | |
WO2010041131A8 (en) | Associating source information with phonetic indices | |
MY165327A (en) | Apparatus,method and computer program for providing one or more adjusted parameters for provision of an upmix signal representation on the basis of a downmix signal representation and a parametric side information associated with the downmix signal representation,using an average value | |
MX2016002561A (es) | Decision sorda/sonora para procesamiento de voz. | |
EP4358085A3 (en) | Signal processing device, method, and program | |
MX355091B (es) | Concepto para codificar una señal de audio y decodificar una señal de audio usando información de conformación espectral relacionada con la voz. | |
MX2012010440A (es) | Procesador de señal, proveedor de ventana, señal de medios codificada, metodo para procesar una señal y metodo para proporcionar una ventana. | |
WO2013132342A3 (en) | Voice signal enhancement | |
MY187944A (en) | Concept for encoding an audio signal and decoding an audio signal using deterministic and noise like information | |
MY172710A (en) | Apparatus and method for generating a frequency enhancement signal using an energy limitation operation | |
MY182138A (en) | Systems and methods of energy-scaled signal processing |