CN1178202C - 用于执行说话者适应或规范化的方法 - Google Patents
用于执行说话者适应或规范化的方法 Download PDFInfo
- Publication number
- CN1178202C CN1178202C CNB991183916A CN99118391A CN1178202C CN 1178202 C CN1178202 C CN 1178202C CN B991183916 A CNB991183916 A CN B991183916A CN 99118391 A CN99118391 A CN 99118391A CN 1178202 C CN1178202 C CN 1178202C
- Authority
- CN
- China
- Prior art keywords
- speaker
- model
- vector
- group
- eigen space
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Expired - Lifetime
Links
- 238000000034 method Methods 0.000 title claims description 134
- 239000013598 vector Substances 0.000 claims abstract description 159
- 238000012549 training Methods 0.000 claims abstract description 89
- 238000007476 Maximum Likelihood Methods 0.000 claims abstract description 49
- 230000006978 adaptation Effects 0.000 claims abstract description 37
- 230000003044 adaptive effect Effects 0.000 claims description 74
- 230000009466 transformation Effects 0.000 claims description 36
- 239000011159 matrix material Substances 0.000 claims description 35
- 230000014509 gene expression Effects 0.000 claims description 30
- 238000006243 chemical reaction Methods 0.000 claims description 26
- 238000009826 distribution Methods 0.000 claims description 25
- 230000008569 process Effects 0.000 claims description 24
- 238000000513 principal component analysis Methods 0.000 claims description 10
- 230000002787 reinforcement Effects 0.000 claims description 10
- 238000000556 factor analysis Methods 0.000 claims description 6
- 238000004458 analytical method Methods 0.000 claims description 5
- 238000000354 decomposition reaction Methods 0.000 claims description 4
- 239000000284 extract Substances 0.000 claims description 4
- 238000012880 independent component analysis Methods 0.000 claims description 4
- 238000005728 strengthening Methods 0.000 claims description 3
- 238000000605 extraction Methods 0.000 claims description 2
- 230000001105 regulatory effect Effects 0.000 claims 2
- 230000004048 modification Effects 0.000 claims 1
- 238000012986 modification Methods 0.000 claims 1
- 230000009467 reduction Effects 0.000 abstract description 15
- 238000013144 data compression Methods 0.000 abstract description 2
- 230000001419 dependent effect Effects 0.000 abstract 1
- 238000005516 engineering process Methods 0.000 description 56
- 230000006870 function Effects 0.000 description 21
- 230000000875 corresponding effect Effects 0.000 description 19
- 230000007613 environmental effect Effects 0.000 description 8
- 230000007704 transition Effects 0.000 description 7
- 238000010586 diagram Methods 0.000 description 6
- 238000012546 transfer Methods 0.000 description 5
- 230000008901 benefit Effects 0.000 description 4
- 238000007667 floating Methods 0.000 description 4
- 241001269238 Data Species 0.000 description 3
- 230000015572 biosynthetic process Effects 0.000 description 2
- 238000005266 casting Methods 0.000 description 2
- 230000002950 deficient Effects 0.000 description 2
- 238000013461 design Methods 0.000 description 2
- 230000008034 disappearance Effects 0.000 description 2
- 230000008676 import Effects 0.000 description 2
- 238000002156 mixing Methods 0.000 description 2
- 239000000203 mixture Substances 0.000 description 2
- 238000012544 monitoring process Methods 0.000 description 2
- 238000005457 optimization Methods 0.000 description 2
- 238000012545 processing Methods 0.000 description 2
- 206010038743 Restlessness Diseases 0.000 description 1
- 238000007792 addition Methods 0.000 description 1
- 238000013459 approach Methods 0.000 description 1
- 150000001875 compounds Chemical class 0.000 description 1
- 230000001143 conditioned effect Effects 0.000 description 1
- 238000010276 construction Methods 0.000 description 1
- 230000002596 correlated effect Effects 0.000 description 1
- 230000004069 differentiation Effects 0.000 description 1
- 230000009977 dual effect Effects 0.000 description 1
- 238000004387 environmental modeling Methods 0.000 description 1
- 238000002474 experimental method Methods 0.000 description 1
- 229910052736 halogen Inorganic materials 0.000 description 1
- 150000002367 halogens Chemical class 0.000 description 1
- 230000002452 interceptive effect Effects 0.000 description 1
- 238000012804 iterative process Methods 0.000 description 1
- 238000013507 mapping Methods 0.000 description 1
- 230000008520 organization Effects 0.000 description 1
- 238000011946 reduction process Methods 0.000 description 1
- 230000004044 response Effects 0.000 description 1
- 230000005236 sound signal Effects 0.000 description 1
Images
Classifications
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L15/00—Speech recognition
- G10L15/06—Creation of reference templates; Training of speech recognition systems, e.g. adaptation to the characteristics of the speaker's voice
- G10L15/065—Adaptation
- G10L15/07—Adaptation to the speaker
-
- G—PHYSICS
- G06—COMPUTING; CALCULATING OR COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F18/00—Pattern recognition
- G06F18/20—Analysing
- G06F18/21—Design or setup of recognition systems or techniques; Extraction of features in feature space; Blind source separation
- G06F18/213—Feature extraction, e.g. by transforming the feature space; Summarisation; Mappings, e.g. subspace methods
- G06F18/2135—Feature extraction, e.g. by transforming the feature space; Summarisation; Mappings, e.g. subspace methods based on approximation criteria, e.g. principal component analysis
Landscapes
- Engineering & Computer Science (AREA)
- Physics & Mathematics (AREA)
- Artificial Intelligence (AREA)
- Theoretical Computer Science (AREA)
- Data Mining & Analysis (AREA)
- Computer Vision & Pattern Recognition (AREA)
- Human Computer Interaction (AREA)
- Bioinformatics & Computational Biology (AREA)
- Multimedia (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Health & Medical Sciences (AREA)
- Life Sciences & Earth Sciences (AREA)
- Bioinformatics & Cheminformatics (AREA)
- Acoustics & Sound (AREA)
- Evolutionary Biology (AREA)
- Evolutionary Computation (AREA)
- General Engineering & Computer Science (AREA)
- General Physics & Mathematics (AREA)
- Computational Linguistics (AREA)
- Stereophonic System (AREA)
- Management, Administration, Business Operations System, And Electronic Commerce (AREA)
- Information Retrieval, Db Structures And Fs Structures Therefor (AREA)
Abstract
Description
Claims (32)
- (修改)1.一种用于执行说话者适应或规范化的方法,该方法包括以下步骤:通过对训练的说话者提供一组模型,来构造表示多个所述训练说话者的本征空间,并对所述模型组执行维数降低,以产生定义所述本征空间的一组基向量;产生适应模型,使用来自新的说话者的输入语音以训练所述适应模型,同时使用所述基向量组来约束所述适应模型,使所述适应模型位于所述本征空间内。
- 2.根据权利要求1的方法,其中通过连接从所述模型组抽取的多个模型参数,并通过对所述模型参数执行线性变换,执行所述维数降低。
- 3.根据权利要求1的方法,其中通过从由主成分分析、线性鉴别分析、因素分析、独立成分分析以及单值分解组成的组中选择的变换过程执行所述维数降低。
- 4.根据权利要求1的方法,其中用于所述训练说话者的模型定义多个模型参数,且所述构造本征空间的步骤包括连接用于所述训练说话者的模型参数以便构造一组超向量,并对所述超向量执行线性维数降低变换,从而产生所述基向量。
- 5.根据权利要求4的方法,其中用于每一所述训练说话者的模型对应于一组不同的语音单元,且其中每一超向量被定义为对应于按预定顺序分类的语音单元的模型参数的连接。
- 6.根据权利要求4的方法,其中所述模型参数为倒谱系数。
- 7.根据权利要求1的方法,其中所述执行维数降低的步骤产生一组数目等于训练说话者数目的基向量。
- 8.根据权利要求1的方法,其中所述执行维数降低的步骤产生基向量的有序列表,并且其中所述构造本征空间的步骤包括放弃所述有序列表的预定部分,以降低所述本征空间阶数。
- 9.根据权利要求1的方法,其中所述约束所述适应模型的步骤通过向所述本征空间投影所述输入语音执行。
- 10.根据权利要求1的方法,其中所述说话者模型组定义多个参数,并且所述方法还包括通过调节所述模型的至少某些参数来强化所述说话者模型,以定义一组强化的说话者模型的步骤。
- 11.根据权利要求10的方法,其中使用极大后验估计执行所述强化步骤。
- 12.根据权利要求10的方法,其中使用基于变换的估计过程执行所述强化步骤。
- 13.根据权利要求10的方法,其中使用极大似然线性回归估计执行所述强化步骤。
- 14.根据权利要求10的方法,其中所述产生适应模型的步骤包括使用来自所述新的说话者的输入语音以产生极大似然向量,以及利用所述极大似然向量来构造所述适应模型,使得所述适应模型位于所述本征空间内。
- 15.根据权利要求1的方法,还包括步骤:通过从所述适应模型抽取模型参数而强化所述适应模型,并基于来自所述新的说话者的输入语音来至少调节某些所述参数。
- 16.根据权利要求15的方法,其中使用极大后验估计执行所述强化步骤。
- 17.根据权利要求15的方法,其中使用基于变换的估计过程执行所述强化步骤。
- 18.根据权利要求15的方法,其中使用极大似然线性回归估计执行所述强化步骤。
- 19.根据权利要求15的方法,其中所述产生适应模型的步骤包括使用来自所述新的说话者的输入语音以产生极大似然向量,以及利用所述极大似然向量来构造所述适应模型,使得所述适应模型位于所述本征空间内。
- 20.根据权利要求19的方法,其中使用极大后验估计执行所述强化步骤。
- 21.根据权利要求19的方法,其中使用基于变换的估计过程执行所述强化步骤。
- 22.根据权利要求19的方法,其中使用极大似然线性回归估计执行所述强化步骤。
- 23.根据权利要求1的方法,其中所述模型组定义第一概率分布,且所述输入语音定义观测数据,且其中所述适应模型的产生使得所述观测数据和所述第一概率分布的乘积最大化。
- 24.根据权利要求23的方法,还包括向所述第一概率分布及所述第二概率分布施加置信因子,以反映由所述分布提供的信息置信度对时间如何变化。
- 25.一种执行说话者适应或规范化的方法,所述方法包括步骤:通过对训练的说话者提供一组模型,构造表示多个所述训练说话者的本征空间,并对所述模型组执行维数降低,以产生定义所述本征空间的一组基向量;产生适应模型,使用来自新的说话者的输入语音以便在定义所述适应模型的本征空间中找出极大似然向量,使所述适应模型位于所述本征空间内。
- 26.根据权利要求25的方法,其中所述产生极大似然向量的步骤包括:定义表示对预定的一组模型产生观测数据的概率的概率函数,其中所述输入语音提供所述观测数据;以及最大化所述概率函数以找出所述极大似然向量。
- 27.根据权利要求25的方法,其中所述适应模型通过使极大似然向量系数乘以所述基向量,而根据所述极大似然向量导出。
- 28.根据权利要求26的方法,其中所述最大化步骤通过以下执行:将所述极大似然向量表示为一组本征值变量;对于所述本征值变量取所述概率函数的一阶导数;以及当所述一阶导数等于零时,求出所述本征值变量对应的值。
- 29.一种执行说话者适应或规范化的方法,该方法包括步骤:将多个训练说话者表示为第一组变换矩阵,以及变换矩阵所适用的模型;通过对所述第一组变换矩阵执行维数降低而构造表示多个训练说话者的本征空间,以产生一组定义所述本征空间的基向量;使用来自新的说话者的输入语音产生第二组变换矩阵,同时使用所述基向量组来约束所述第二组变换矩阵,使得所述第二组变换矩阵位于所述本征空间内。
- 30.根据权利要求29的方法,其中所述第一组变换矩阵是通过极大似然线性回归产生的。
- 31.根据权利要求29的方法,还包括使所述第一组变换矩阵每一个向量化以定义一组超向量,并对所述超向量执行维数降低以定义所述本征空间。
- 32.根据权利要求29的方法,还包括使用来自新说话者的输入语音产生所述第二组变换矩阵,以产生极大似然向量,使用所述极大似然向量确定所述本征空间内的位置。
Applications Claiming Priority (2)
Application Number | Priority Date | Filing Date | Title |
---|---|---|---|
US09/148,753 | 1998-09-04 | ||
US09/148,753 US6343267B1 (en) | 1998-04-30 | 1998-09-04 | Dimensionality reduction for speaker normalization and speaker and environment adaptation using eigenvoice techniques |
Publications (2)
Publication Number | Publication Date |
---|---|
CN1253353A CN1253353A (zh) | 2000-05-17 |
CN1178202C true CN1178202C (zh) | 2004-12-01 |
Family
ID=22527202
Family Applications (1)
Application Number | Title | Priority Date | Filing Date |
---|---|---|---|
CNB991183916A Expired - Lifetime CN1178202C (zh) | 1998-09-04 | 1999-09-03 | 用于执行说话者适应或规范化的方法 |
Country Status (6)
Country | Link |
---|---|
US (1) | US6343267B1 (zh) |
EP (1) | EP0984429B1 (zh) |
JP (1) | JP2000081893A (zh) |
CN (1) | CN1178202C (zh) |
DE (1) | DE69916951T2 (zh) |
TW (1) | TW452758B (zh) |
Cited By (3)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
US7658794B2 (en) | 2000-03-14 | 2010-02-09 | James Hardie Technology Limited | Fiber cement building materials with low density additives |
US7704316B2 (en) | 2001-03-02 | 2010-04-27 | James Hardie Technology Limited | Coatings for building products and methods of making same |
US8209927B2 (en) | 2007-12-20 | 2012-07-03 | James Hardie Technology Limited | Structural fiber cement building materials |
Families Citing this family (183)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
US6343267B1 (en) | 1998-04-30 | 2002-01-29 | Matsushita Electric Industrial Co., Ltd. | Dimensionality reduction for speaker normalization and speaker and environment adaptation using eigenvoice techniques |
US8095581B2 (en) * | 1999-02-05 | 2012-01-10 | Gregory A Stobbs | Computer-implemented patent portfolio analysis method and apparatus |
EP1159734B1 (de) * | 1999-03-08 | 2004-05-19 | Siemens Aktiengesellschaft | Verfahren und anordnung zur ermittlung einer merkmalsbeschreibung eines sprachsignals |
KR100307623B1 (ko) * | 1999-10-21 | 2001-11-02 | 윤종용 | 엠.에이.피 화자 적응 조건에서 파라미터의 분별적 추정 방법 및 장치 및 이를 각각 포함한 음성 인식 방법 및 장치 |
US6571208B1 (en) * | 1999-11-29 | 2003-05-27 | Matsushita Electric Industrial Co., Ltd. | Context-dependent acoustic models for medium and large vocabulary speech recognition with eigenvoice training |
US6526379B1 (en) * | 1999-11-29 | 2003-02-25 | Matsushita Electric Industrial Co., Ltd. | Discriminative clustering methods for automatic speech recognition |
US6868381B1 (en) * | 1999-12-21 | 2005-03-15 | Nortel Networks Limited | Method and apparatus providing hypothesis driven speech modelling for use in speech recognition |
JP5105682B2 (ja) * | 2000-02-25 | 2012-12-26 | ニュアンス コミュニケーションズ オーストリア ゲーエムベーハー | 基準変換手段を伴なう音声認識装置 |
US8645137B2 (en) | 2000-03-16 | 2014-02-04 | Apple Inc. | Fast, language-independent method for user authentication by voice |
US20080147404A1 (en) * | 2000-05-15 | 2008-06-19 | Nusuara Technologies Sdn Bhd | System and methods for accent classification and adaptation |
US6751590B1 (en) * | 2000-06-13 | 2004-06-15 | International Business Machines Corporation | Method and apparatus for performing pattern-specific maximum likelihood transformations for speaker recognition |
US6961703B1 (en) * | 2000-09-13 | 2005-11-01 | Itt Manufacturing Enterprises, Inc. | Method for speech processing involving whole-utterance modeling |
DE10047724A1 (de) * | 2000-09-27 | 2002-04-11 | Philips Corp Intellectual Pty | Verfahren zur Ermittlung eines Eigenraumes zur Darstellung einer Mehrzahl von Trainingssprechern |
DE10047718A1 (de) * | 2000-09-27 | 2002-04-18 | Philips Corp Intellectual Pty | Verfahren zur Spracherkennung |
DE10047723A1 (de) * | 2000-09-27 | 2002-04-11 | Philips Corp Intellectual Pty | Verfahren zur Ermittlung eines Eigenraums zur Darstellung einer Mehrzahl von Trainingssprechern |
US7006969B2 (en) * | 2000-11-02 | 2006-02-28 | At&T Corp. | System and method of pattern recognition in very high-dimensional space |
US6895376B2 (en) * | 2001-05-04 | 2005-05-17 | Matsushita Electric Industrial Co., Ltd. | Eigenvoice re-estimation technique of acoustic models for speech recognition, speaker identification and speaker verification |
JP2002366187A (ja) * | 2001-06-08 | 2002-12-20 | Sony Corp | 音声認識装置および音声認識方法、並びにプログラムおよび記録媒体 |
US7050969B2 (en) * | 2001-11-27 | 2006-05-23 | Mitsubishi Electric Research Laboratories, Inc. | Distributed speech recognition with codec parameters |
US7209881B2 (en) | 2001-12-20 | 2007-04-24 | Matsushita Electric Industrial Co., Ltd. | Preparing acoustic models by sufficient statistics and noise-superimposed speech data |
US7472062B2 (en) * | 2002-01-04 | 2008-12-30 | International Business Machines Corporation | Efficient recursive clustering based on a splitting function derived from successive eigen-decompositions |
US7139703B2 (en) * | 2002-04-05 | 2006-11-21 | Microsoft Corporation | Method of iterative noise estimation in a recursive framework |
US20030195751A1 (en) * | 2002-04-10 | 2003-10-16 | Mitsubishi Electric Research Laboratories, Inc. | Distributed automatic speech recognition with persistent user parameters |
US20040117181A1 (en) * | 2002-09-24 | 2004-06-17 | Keiko Morii | Method of speaker normalization for speech recognition using frequency conversion and speech recognition apparatus applying the preceding method |
US20040122672A1 (en) * | 2002-12-18 | 2004-06-24 | Jean-Francois Bonastre | Gaussian model-based dynamic time warping system and method for speech processing |
US7165026B2 (en) | 2003-03-31 | 2007-01-16 | Microsoft Corporation | Method of noise estimation using incremental bayes learning |
US7516157B2 (en) * | 2003-05-08 | 2009-04-07 | Microsoft Corporation | Relational directory |
US8229744B2 (en) * | 2003-08-26 | 2012-07-24 | Nuance Communications, Inc. | Class detection scheme and time mediated averaging of class dependent models |
US20080208581A1 (en) * | 2003-12-05 | 2008-08-28 | Queensland University Of Technology | Model Adaptation System and Method for Speaker Recognition |
KR100612840B1 (ko) * | 2004-02-18 | 2006-08-18 | 삼성전자주식회사 | 모델 변이 기반의 화자 클러스터링 방법, 화자 적응 방법및 이들을 이용한 음성 인식 장치 |
GB2414328A (en) * | 2004-05-17 | 2005-11-23 | Mitsubishi Electric Inf Tech | Discrimination transforms applied to frequency domain derived feature vectors |
US7496509B2 (en) * | 2004-05-28 | 2009-02-24 | International Business Machines Corporation | Methods and apparatus for statistical biometric model migration |
US7567903B1 (en) * | 2005-01-12 | 2009-07-28 | At&T Intellectual Property Ii, L.P. | Low latency real-time vocal tract length normalization |
WO2006076661A2 (en) * | 2005-01-14 | 2006-07-20 | Tremor Media Llc | Dynamic advertisement system and method |
US20070049367A1 (en) * | 2005-08-23 | 2007-03-01 | Way Out World, Llc | Methods for game augmented interactive marketing |
US20070050243A1 (en) * | 2005-08-23 | 2007-03-01 | Way Out World, Llc | Multi-unit system and methods for game augmented interactive marketing |
US20070050242A1 (en) * | 2005-08-23 | 2007-03-01 | Way Out World, Llc | Solo-unit system and methods for game augmented interactive marketing |
US8677377B2 (en) | 2005-09-08 | 2014-03-18 | Apple Inc. | Method and apparatus for building an intelligent automated assistant |
JP2007114413A (ja) * | 2005-10-19 | 2007-05-10 | Toshiba Corp | 音声非音声判別装置、音声区間検出装置、音声非音声判別方法、音声区間検出方法、音声非音声判別プログラムおよび音声区間検出プログラム |
JP2009521736A (ja) * | 2005-11-07 | 2009-06-04 | スキャンスカウト,インコーポレイテッド | リッチメディアと共に広告をレンダリングするための技術 |
US9318108B2 (en) | 2010-01-18 | 2016-04-19 | Apple Inc. | Intelligent automated assistant |
JP4282704B2 (ja) * | 2006-09-27 | 2009-06-24 | 株式会社東芝 | 音声区間検出装置およびプログラム |
US20080109391A1 (en) * | 2006-11-07 | 2008-05-08 | Scanscout, Inc. | Classifying content based on mood |
US8977255B2 (en) | 2007-04-03 | 2015-03-10 | Apple Inc. | Method and system for operating a multi-function portable electronic device using voice-activation |
JP2010539085A (ja) * | 2007-09-07 | 2010-12-16 | バイオノボ・インコーポレーテッド | マメ科ファミリーのキバナオウギのエストロゲン性抽出物およびその使用 |
US8549550B2 (en) | 2008-09-17 | 2013-10-01 | Tubemogul, Inc. | Method and apparatus for passively monitoring online video viewing and viewer behavior |
US8577996B2 (en) * | 2007-09-18 | 2013-11-05 | Tremor Video, Inc. | Method and apparatus for tracing users of online video web sites |
US9330720B2 (en) | 2008-01-03 | 2016-05-03 | Apple Inc. | Methods and apparatus for altering audio output signals |
US8775416B2 (en) * | 2008-01-09 | 2014-07-08 | Yahoo!Inc. | Adapting a context-independent relevance function for identifying relevant search results |
JP4950930B2 (ja) * | 2008-04-03 | 2012-06-13 | 株式会社東芝 | 音声/非音声を判定する装置、方法およびプログラム |
US8996376B2 (en) | 2008-04-05 | 2015-03-31 | Apple Inc. | Intelligent text-to-speech conversion |
US20090259552A1 (en) * | 2008-04-11 | 2009-10-15 | Tremor Media, Inc. | System and method for providing advertisements from multiple ad servers using a failover mechanism |
US10496753B2 (en) | 2010-01-18 | 2019-12-03 | Apple Inc. | Automatically adapting user interfaces for hands-free interaction |
US8521530B1 (en) | 2008-06-30 | 2013-08-27 | Audience, Inc. | System and method for enhancing a monaural audio signal |
US20100030549A1 (en) | 2008-07-31 | 2010-02-04 | Lee Michael M | Mobile device having human language translation capability with positional feedback |
US9612995B2 (en) | 2008-09-17 | 2017-04-04 | Adobe Systems Incorporated | Video viewer targeting based on preference similarity |
WO2010067118A1 (en) | 2008-12-11 | 2010-06-17 | Novauris Technologies Limited | Speech recognition involving a mobile device |
US20120309363A1 (en) | 2011-06-03 | 2012-12-06 | Apple Inc. | Triggering notifications associated with tasks items that represent tasks to perform |
US10241644B2 (en) | 2011-06-03 | 2019-03-26 | Apple Inc. | Actionable reminder entries |
US9858925B2 (en) | 2009-06-05 | 2018-01-02 | Apple Inc. | Using context information to facilitate processing of commands in a virtual assistant |
US10241752B2 (en) | 2011-09-30 | 2019-03-26 | Apple Inc. | Interface for a virtual digital assistant |
US9431006B2 (en) | 2009-07-02 | 2016-08-30 | Apple Inc. | Methods and apparatuses for automatic speech recognition |
US20110093783A1 (en) * | 2009-10-16 | 2011-04-21 | Charles Parra | Method and system for linking media components |
US8374867B2 (en) * | 2009-11-13 | 2013-02-12 | At&T Intellectual Property I, L.P. | System and method for standardized speech recognition infrastructure |
EP2502195A2 (en) * | 2009-11-20 | 2012-09-26 | Tadashi Yonezaki | Methods and apparatus for optimizing advertisement allocation |
US10276170B2 (en) | 2010-01-18 | 2019-04-30 | Apple Inc. | Intelligent automated assistant |
US10553209B2 (en) | 2010-01-18 | 2020-02-04 | Apple Inc. | Systems and methods for hands-free notification summaries |
US10679605B2 (en) | 2010-01-18 | 2020-06-09 | Apple Inc. | Hands-free list-reading by intelligent automated assistant |
US10705794B2 (en) | 2010-01-18 | 2020-07-07 | Apple Inc. | Automatically adapting user interfaces for hands-free interaction |
DE202011111062U1 (de) | 2010-01-25 | 2019-02-19 | Newvaluexchange Ltd. | Vorrichtung und System für eine Digitalkonversationsmanagementplattform |
US9008329B1 (en) * | 2010-01-26 | 2015-04-14 | Audience, Inc. | Noise reduction using multi-feature cluster tracker |
US8682667B2 (en) | 2010-02-25 | 2014-03-25 | Apple Inc. | User profiling for selecting user specific voice input processing information |
US8473287B2 (en) | 2010-04-19 | 2013-06-25 | Audience, Inc. | Method for jointly optimizing noise reduction and voice quality in a mono or multi-microphone system |
US8538035B2 (en) | 2010-04-29 | 2013-09-17 | Audience, Inc. | Multi-microphone robust noise suppression |
US8781137B1 (en) | 2010-04-27 | 2014-07-15 | Audience, Inc. | Wind noise detection and suppression |
US9558755B1 (en) | 2010-05-20 | 2017-01-31 | Knowles Electronics, Llc | Noise suppression assisted automatic speech recognition |
US8447596B2 (en) | 2010-07-12 | 2013-05-21 | Audience, Inc. | Monaural noise suppression based on computational auditory scene analysis |
US10762293B2 (en) | 2010-12-22 | 2020-09-01 | Apple Inc. | Using parts-of-speech tagging and named entity recognition for spelling correction |
US9262612B2 (en) | 2011-03-21 | 2016-02-16 | Apple Inc. | Device access using voice authentication |
US10057736B2 (en) | 2011-06-03 | 2018-08-21 | Apple Inc. | Active transport based notifications |
US8994660B2 (en) | 2011-08-29 | 2015-03-31 | Apple Inc. | Text correction processing |
US10134385B2 (en) | 2012-03-02 | 2018-11-20 | Apple Inc. | Systems and methods for name pronunciation |
US9483461B2 (en) | 2012-03-06 | 2016-11-01 | Apple Inc. | Handling speech synthesis of content for multiple languages |
US9406299B2 (en) * | 2012-05-08 | 2016-08-02 | Nuance Communications, Inc. | Differential acoustic model representation and linear transform-based adaptation for efficient user profile update techniques in automatic speech recognition |
US9280610B2 (en) | 2012-05-14 | 2016-03-08 | Apple Inc. | Crowd sourcing information to fulfill user requests |
US8972312B2 (en) | 2012-05-29 | 2015-03-03 | Nuance Communications, Inc. | Methods and apparatus for performing transformation techniques for data clustering and/or classification |
US9721563B2 (en) | 2012-06-08 | 2017-08-01 | Apple Inc. | Name recognition system |
US9495129B2 (en) | 2012-06-29 | 2016-11-15 | Apple Inc. | Device, method, and user interface for voice-activated navigation and browsing of a document |
US9679556B2 (en) * | 2012-08-24 | 2017-06-13 | Interactive Intelligence Group, Inc. | Method and system for selectively biased linear discriminant analysis in automatic speech recognition systems |
US9576574B2 (en) | 2012-09-10 | 2017-02-21 | Apple Inc. | Context-sensitive handling of interruptions by intelligent digital assistant |
US9547647B2 (en) | 2012-09-19 | 2017-01-17 | Apple Inc. | Voice-based media searching |
US9640194B1 (en) | 2012-10-04 | 2017-05-02 | Knowles Electronics, Llc | Noise suppression for speech processing based on machine-learning mask estimation |
GB2510200B (en) * | 2013-01-29 | 2017-05-10 | Toshiba Res Europe Ltd | A computer generated head |
CN113470640B (zh) | 2013-02-07 | 2022-04-26 | 苹果公司 | 数字助理的语音触发器 |
US9368114B2 (en) | 2013-03-14 | 2016-06-14 | Apple Inc. | Context-sensitive handling of interruptions |
WO2014144949A2 (en) | 2013-03-15 | 2014-09-18 | Apple Inc. | Training an at least partial voice command system |
WO2014144579A1 (en) | 2013-03-15 | 2014-09-18 | Apple Inc. | System and method for updating an adaptive speech recognition model |
WO2014197334A2 (en) | 2013-06-07 | 2014-12-11 | Apple Inc. | System and method for user-specified pronunciation of words for speech synthesis and recognition |
US9582608B2 (en) | 2013-06-07 | 2017-02-28 | Apple Inc. | Unified ranking with entropy-weighted information for phrase-based semantic auto-completion |
WO2014197336A1 (en) | 2013-06-07 | 2014-12-11 | Apple Inc. | System and method for detecting errors in interactions with a voice-based digital assistant |
WO2014197335A1 (en) | 2013-06-08 | 2014-12-11 | Apple Inc. | Interpreting and acting upon commands that involve sharing information with remote devices |
WO2014200728A1 (en) | 2013-06-09 | 2014-12-18 | Apple Inc. | Device, method, and graphical user interface for enabling conversation persistence across two or more instances of a digital assistant |
US10176167B2 (en) | 2013-06-09 | 2019-01-08 | Apple Inc. | System and method for inferring user intent from speech inputs |
AU2014278595B2 (en) | 2013-06-13 | 2017-04-06 | Apple Inc. | System and method for emergency calls initiated by voice command |
WO2015020942A1 (en) | 2013-08-06 | 2015-02-12 | Apple Inc. | Auto-activating smart responses based on activities from remote devices |
US9251784B2 (en) | 2013-10-23 | 2016-02-02 | International Business Machines Corporation | Regularized feature space discrimination adaptation |
CN106462772B (zh) | 2014-02-19 | 2019-12-13 | 河谷控股Ip有限责任公司 | 对象识别特征的基于不变量的维数缩减、系统和方法 |
US9620105B2 (en) | 2014-05-15 | 2017-04-11 | Apple Inc. | Analyzing audio input for efficient speech and music recognition |
US10592095B2 (en) | 2014-05-23 | 2020-03-17 | Apple Inc. | Instantaneous speaking of content on touch devices |
US9502031B2 (en) | 2014-05-27 | 2016-11-22 | Apple Inc. | Method for supporting dynamic grammars in WFST-based ASR |
US9633004B2 (en) | 2014-05-30 | 2017-04-25 | Apple Inc. | Better resolution when referencing to concepts |
US10078631B2 (en) | 2014-05-30 | 2018-09-18 | Apple Inc. | Entropy-guided text prediction using combined word and character n-gram language models |
US9734193B2 (en) | 2014-05-30 | 2017-08-15 | Apple Inc. | Determining domain salience ranking from ambiguous words in natural speech |
US10289433B2 (en) | 2014-05-30 | 2019-05-14 | Apple Inc. | Domain specific language for encoding assistant dialog |
US9760559B2 (en) | 2014-05-30 | 2017-09-12 | Apple Inc. | Predictive text input |
US9715875B2 (en) | 2014-05-30 | 2017-07-25 | Apple Inc. | Reducing the need for manual start/end-pointing and trigger phrases |
US9842101B2 (en) | 2014-05-30 | 2017-12-12 | Apple Inc. | Predictive conversion of language input |
EP3480811A1 (en) | 2014-05-30 | 2019-05-08 | Apple Inc. | Multi-command single utterance input method |
US10170123B2 (en) | 2014-05-30 | 2019-01-01 | Apple Inc. | Intelligent assistant for home automation |
US9785630B2 (en) | 2014-05-30 | 2017-10-10 | Apple Inc. | Text prediction using combined word N-gram and unigram language models |
US9430463B2 (en) | 2014-05-30 | 2016-08-30 | Apple Inc. | Exemplar-based natural language processing |
US10659851B2 (en) | 2014-06-30 | 2020-05-19 | Apple Inc. | Real-time digital assistant knowledge updates |
US9338493B2 (en) | 2014-06-30 | 2016-05-10 | Apple Inc. | Intelligent automated assistant for TV user interactions |
US10446141B2 (en) | 2014-08-28 | 2019-10-15 | Apple Inc. | Automatic speech recognition based on user feedback |
WO2016033364A1 (en) | 2014-08-28 | 2016-03-03 | Audience, Inc. | Multi-sourced noise suppression |
US9818400B2 (en) | 2014-09-11 | 2017-11-14 | Apple Inc. | Method and apparatus for discovering trending terms in speech requests |
US10789041B2 (en) | 2014-09-12 | 2020-09-29 | Apple Inc. | Dynamic thresholds for always listening speech trigger |
US9606986B2 (en) | 2014-09-29 | 2017-03-28 | Apple Inc. | Integrated word N-gram and class M-gram language models |
US9646609B2 (en) | 2014-09-30 | 2017-05-09 | Apple Inc. | Caching apparatus for serving phonetic pronunciations |
US9886432B2 (en) | 2014-09-30 | 2018-02-06 | Apple Inc. | Parsimonious handling of word inflection via categorical stem + suffix N-gram language models |
US10074360B2 (en) | 2014-09-30 | 2018-09-11 | Apple Inc. | Providing an indication of the suitability of speech recognition |
US10127911B2 (en) | 2014-09-30 | 2018-11-13 | Apple Inc. | Speaker identification and unsupervised speaker adaptation techniques |
US9668121B2 (en) | 2014-09-30 | 2017-05-30 | Apple Inc. | Social reminders |
US10552013B2 (en) | 2014-12-02 | 2020-02-04 | Apple Inc. | Data detection |
US9711141B2 (en) | 2014-12-09 | 2017-07-18 | Apple Inc. | Disambiguating heteronyms in speech synthesis |
US9865280B2 (en) | 2015-03-06 | 2018-01-09 | Apple Inc. | Structured dictation using intelligent automated assistants |
US10567477B2 (en) | 2015-03-08 | 2020-02-18 | Apple Inc. | Virtual assistant continuity |
US9886953B2 (en) | 2015-03-08 | 2018-02-06 | Apple Inc. | Virtual assistant activation |
US9721566B2 (en) | 2015-03-08 | 2017-08-01 | Apple Inc. | Competing devices responding to voice triggers |
US9899019B2 (en) | 2015-03-18 | 2018-02-20 | Apple Inc. | Systems and methods for structured stem and suffix language models |
US9842105B2 (en) | 2015-04-16 | 2017-12-12 | Apple Inc. | Parsimonious continuous-space phrase representations for natural language processing |
US10083688B2 (en) | 2015-05-27 | 2018-09-25 | Apple Inc. | Device voice control for selecting a displayed affordance |
US10127220B2 (en) | 2015-06-04 | 2018-11-13 | Apple Inc. | Language identification from short strings |
US9578173B2 (en) | 2015-06-05 | 2017-02-21 | Apple Inc. | Virtual assistant aided communication with 3rd party service in a communication session |
US10101822B2 (en) | 2015-06-05 | 2018-10-16 | Apple Inc. | Language input correction |
US10186254B2 (en) | 2015-06-07 | 2019-01-22 | Apple Inc. | Context-based endpoint detection |
US11025565B2 (en) | 2015-06-07 | 2021-06-01 | Apple Inc. | Personalized prediction of responses for instant messaging |
US10255907B2 (en) | 2015-06-07 | 2019-04-09 | Apple Inc. | Automatic accent detection using acoustic models |
US10671428B2 (en) | 2015-09-08 | 2020-06-02 | Apple Inc. | Distributed personal assistant |
US10747498B2 (en) | 2015-09-08 | 2020-08-18 | Apple Inc. | Zero latency digital assistant |
US9697820B2 (en) | 2015-09-24 | 2017-07-04 | Apple Inc. | Unit-selection text-to-speech synthesis using concatenation-sensitive neural networks |
US10366158B2 (en) | 2015-09-29 | 2019-07-30 | Apple Inc. | Efficient word encoding for recurrent neural network language models |
US11010550B2 (en) | 2015-09-29 | 2021-05-18 | Apple Inc. | Unified language modeling framework for word prediction, auto-completion and auto-correction |
US11587559B2 (en) | 2015-09-30 | 2023-02-21 | Apple Inc. | Intelligent device identification |
US10691473B2 (en) | 2015-11-06 | 2020-06-23 | Apple Inc. | Intelligent automated assistant in a messaging environment |
US10049668B2 (en) | 2015-12-02 | 2018-08-14 | Apple Inc. | Applying neural network language models to weighted finite state transducers for automatic speech recognition |
US10223066B2 (en) | 2015-12-23 | 2019-03-05 | Apple Inc. | Proactive assistance based on dialog communication between devices |
US10446143B2 (en) | 2016-03-14 | 2019-10-15 | Apple Inc. | Identification of voice inputs providing credentials |
US9934775B2 (en) | 2016-05-26 | 2018-04-03 | Apple Inc. | Unit-selection text-to-speech synthesis based on predicted concatenation parameters |
US9972304B2 (en) | 2016-06-03 | 2018-05-15 | Apple Inc. | Privacy preserving distributed evaluation framework for embedded personalized systems |
US10249300B2 (en) | 2016-06-06 | 2019-04-02 | Apple Inc. | Intelligent list reading |
US10049663B2 (en) | 2016-06-08 | 2018-08-14 | Apple, Inc. | Intelligent automated assistant for media exploration |
DK179309B1 (en) | 2016-06-09 | 2018-04-23 | Apple Inc | Intelligent automated assistant in a home environment |
US10509862B2 (en) | 2016-06-10 | 2019-12-17 | Apple Inc. | Dynamic phrase expansion of language input |
US10490187B2 (en) | 2016-06-10 | 2019-11-26 | Apple Inc. | Digital assistant providing automated status report |
US10586535B2 (en) | 2016-06-10 | 2020-03-10 | Apple Inc. | Intelligent digital assistant in a multi-tasking environment |
US10067938B2 (en) | 2016-06-10 | 2018-09-04 | Apple Inc. | Multilingual word prediction |
US10192552B2 (en) | 2016-06-10 | 2019-01-29 | Apple Inc. | Digital assistant providing whispered speech |
DK201670540A1 (en) | 2016-06-11 | 2018-01-08 | Apple Inc | Application integration with a digital assistant |
DK179049B1 (en) | 2016-06-11 | 2017-09-18 | Apple Inc | Data driven natural language event detection and classification |
DK179343B1 (en) | 2016-06-11 | 2018-05-14 | Apple Inc | Intelligent task discovery |
DK179415B1 (en) | 2016-06-11 | 2018-06-14 | Apple Inc | Intelligent device arbitration and control |
US10043516B2 (en) | 2016-09-23 | 2018-08-07 | Apple Inc. | Intelligent automated assistant |
US10593346B2 (en) | 2016-12-22 | 2020-03-17 | Apple Inc. | Rank-reduced token representation for automatic speech recognition |
DK201770439A1 (en) | 2017-05-11 | 2018-12-13 | Apple Inc. | Offline personal assistant |
DK179745B1 (en) | 2017-05-12 | 2019-05-01 | Apple Inc. | SYNCHRONIZATION AND TASK DELEGATION OF A DIGITAL ASSISTANT |
DK179496B1 (en) | 2017-05-12 | 2019-01-15 | Apple Inc. | USER-SPECIFIC Acoustic Models |
DK201770432A1 (en) | 2017-05-15 | 2018-12-21 | Apple Inc. | Hierarchical belief states for digital assistants |
DK201770431A1 (en) | 2017-05-15 | 2018-12-20 | Apple Inc. | Optimizing dialogue policy decisions for digital assistants using implicit feedback |
DK179560B1 (en) | 2017-05-16 | 2019-02-18 | Apple Inc. | FAR-FIELD EXTENSION FOR DIGITAL ASSISTANT SERVICES |
EP3553775B1 (en) * | 2018-04-12 | 2020-11-25 | Spotify AB | Voice-based authentication |
CN109887484B (zh) * | 2019-02-22 | 2023-08-04 | 平安科技(深圳)有限公司 | 一种基于对偶学习的语音识别与语音合成方法及装置 |
CN113505801B (zh) * | 2021-09-13 | 2021-11-30 | 拓小拓科技(天津)有限公司 | 一种用于超维计算的图像编码方法 |
Family Cites Families (35)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
US4903035A (en) | 1983-12-20 | 1990-02-20 | Bsh Electronics, Ltd. | Electrical signal separating device having isolating and matching circuitry |
US4718088A (en) | 1984-03-27 | 1988-01-05 | Exxon Research And Engineering Company | Speech recognition training method |
JPS62231993A (ja) | 1986-03-25 | 1987-10-12 | インタ−ナシヨナル ビジネス マシ−ンズ コ−ポレ−シヨン | 音声認識方法 |
US4817156A (en) | 1987-08-10 | 1989-03-28 | International Business Machines Corporation | Rapidly training a speech recognizer to a subsequent speaker given training data of a reference speaker |
JPH01102599A (ja) | 1987-10-12 | 1989-04-20 | Internatl Business Mach Corp <Ibm> | 音声認識方法 |
JP2733955B2 (ja) | 1988-05-18 | 1998-03-30 | 日本電気株式会社 | 適応型音声認識装置 |
US5127055A (en) | 1988-12-30 | 1992-06-30 | Kurzweil Applied Intelligence, Inc. | Speech recognition apparatus & method having dynamic reference pattern adaptation |
JPH0636156B2 (ja) | 1989-03-13 | 1994-05-11 | インターナショナル・ビジネス・マシーンズ・コーポレーション | 音声認識装置 |
DE3931638A1 (de) | 1989-09-22 | 1991-04-04 | Standard Elektrik Lorenz Ag | Verfahren zur sprecheradaptiven erkennung von sprache |
JP3014177B2 (ja) | 1991-08-08 | 2000-02-28 | 富士通株式会社 | 話者適応音声認識装置 |
US5280562A (en) * | 1991-10-03 | 1994-01-18 | International Business Machines Corporation | Speech coding apparatus with single-dimension acoustic prototypes for a speech recognizer |
US5278942A (en) * | 1991-12-05 | 1994-01-11 | International Business Machines Corporation | Speech coding apparatus having speaker dependent prototypes generated from nonuser reference data |
DE69322894T2 (de) | 1992-03-02 | 1999-07-29 | At & T Corp | Lernverfahren und Gerät zur Spracherkennung |
US5233681A (en) * | 1992-04-24 | 1993-08-03 | International Business Machines Corporation | Context-dependent speech recognizer using estimated next word context |
US5293584A (en) * | 1992-05-21 | 1994-03-08 | International Business Machines Corporation | Speech recognition system for natural language translation |
US5473728A (en) | 1993-02-24 | 1995-12-05 | The United States Of America As Represented By The Secretary Of The Navy | Training of homoscedastic hidden Markov models for automatic speech recognition |
JPH075892A (ja) | 1993-04-29 | 1995-01-10 | Matsushita Electric Ind Co Ltd | 音声認識方法 |
US5664059A (en) | 1993-04-29 | 1997-09-02 | Panasonic Technologies, Inc. | Self-learning speaker adaptation based on spectral variation source decomposition |
US5522011A (en) * | 1993-09-27 | 1996-05-28 | International Business Machines Corporation | Speech coding apparatus and method using classification rules |
WO1995009416A1 (en) | 1993-09-30 | 1995-04-06 | Apple Computer, Inc. | Continuous reference adaptation in a pattern recognition system |
JP2692581B2 (ja) | 1994-06-07 | 1997-12-17 | 日本電気株式会社 | 音響カテゴリ平均値計算装置及び適応化装置 |
US5793891A (en) | 1994-07-07 | 1998-08-11 | Nippon Telegraph And Telephone Corporation | Adaptive training method for pattern recognition |
US5825978A (en) * | 1994-07-18 | 1998-10-20 | Sri International | Method and apparatus for speech recognition using optimized partial mixture tying of HMM state functions |
US5737723A (en) | 1994-08-29 | 1998-04-07 | Lucent Technologies Inc. | Confusable word detection in speech recognition |
US5864810A (en) * | 1995-01-20 | 1999-01-26 | Sri International | Method and apparatus for speech recognition adapted to an individual speaker |
JP3453456B2 (ja) | 1995-06-19 | 2003-10-06 | キヤノン株式会社 | 状態共有モデルの設計方法及び装置ならびにその状態共有モデルを用いた音声認識方法および装置 |
US5842163A (en) * | 1995-06-21 | 1998-11-24 | Sri International | Method and apparatus for computing likelihood and hypothesizing keyword appearance in speech |
US5806029A (en) | 1995-09-15 | 1998-09-08 | At&T Corp | Signal conditioned minimum error rate training for continuous speech recognition |
JP2871561B2 (ja) | 1995-11-30 | 1999-03-17 | 株式会社エイ・ティ・アール音声翻訳通信研究所 | 不特定話者モデル生成装置及び音声認識装置 |
US5787394A (en) | 1995-12-13 | 1998-07-28 | International Business Machines Corporation | State-dependent speaker clustering for speaker adaptation |
US5778342A (en) | 1996-02-01 | 1998-07-07 | Dspc Israel Ltd. | Pattern recognition system and method |
US5895447A (en) | 1996-02-02 | 1999-04-20 | International Business Machines Corporation | Speech recognition using thresholded speaker class model selection or model adaptation |
JP3302266B2 (ja) | 1996-07-23 | 2002-07-15 | 沖電気工業株式会社 | ヒドン・マルコフ・モデルの学習方法 |
US6343267B1 (en) | 1998-04-30 | 2002-01-29 | Matsushita Electric Industrial Co., Ltd. | Dimensionality reduction for speaker normalization and speaker and environment adaptation using eigenvoice techniques |
TW436758B (en) | 1998-04-30 | 2001-05-28 | Matsushita Electric Ind Co Ltd | Speaker and environment adaptation based on eigenvoices including maximum likelihood method |
-
1998
- 1998-09-04 US US09/148,753 patent/US6343267B1/en not_active Expired - Lifetime
-
1999
- 1999-08-23 EP EP99306667A patent/EP0984429B1/en not_active Expired - Lifetime
- 1999-08-23 DE DE69916951T patent/DE69916951T2/de not_active Expired - Fee Related
- 1999-09-03 CN CNB991183916A patent/CN1178202C/zh not_active Expired - Lifetime
- 1999-09-06 JP JP11251083A patent/JP2000081893A/ja active Pending
- 1999-11-09 TW TW088115209A patent/TW452758B/zh not_active IP Right Cessation
Cited By (6)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
US7658794B2 (en) | 2000-03-14 | 2010-02-09 | James Hardie Technology Limited | Fiber cement building materials with low density additives |
US7727329B2 (en) | 2000-03-14 | 2010-06-01 | James Hardie Technology Limited | Fiber cement building materials with low density additives |
US8182606B2 (en) | 2000-03-14 | 2012-05-22 | James Hardie Technology Limited | Fiber cement building materials with low density additives |
US8603239B2 (en) | 2000-03-14 | 2013-12-10 | James Hardie Technology Limited | Fiber cement building materials with low density additives |
US7704316B2 (en) | 2001-03-02 | 2010-04-27 | James Hardie Technology Limited | Coatings for building products and methods of making same |
US8209927B2 (en) | 2007-12-20 | 2012-07-03 | James Hardie Technology Limited | Structural fiber cement building materials |
Also Published As
Publication number | Publication date |
---|---|
TW452758B (en) | 2001-09-01 |
DE69916951D1 (de) | 2004-06-09 |
EP0984429A2 (en) | 2000-03-08 |
CN1253353A (zh) | 2000-05-17 |
DE69916951T2 (de) | 2005-06-23 |
US6343267B1 (en) | 2002-01-29 |
EP0984429A3 (en) | 2000-11-22 |
EP0984429B1 (en) | 2004-05-06 |
JP2000081893A (ja) | 2000-03-21 |
Similar Documents
Publication | Publication Date | Title |
---|---|---|
CN1178202C (zh) | 用于执行说话者适应或规范化的方法 | |
CN1188828C (zh) | 基于本征话音的说话者检验和说话者识别 | |
CN100347741C (zh) | 移动语音合成方法 | |
Tam et al. | Dynamic language model adaptation using variational Bayes inference. | |
CN1229773C (zh) | 语音识别对话装置 | |
CN1150515C (zh) | 语音识别方法和装置 | |
CN1234109C (zh) | 语调生成方法、语音合成装置、语音合成方法及语音服务器 | |
CN1591570A (zh) | 用于紧凑声学建模的泡分裂法 | |
CN101051215A (zh) | 学习设备、学习方法和程序 | |
CN101079266A (zh) | 基于多统计模型和最小均方误差实现背景噪声抑制的方法 | |
CN1534597A (zh) | 利用具有转换状态空间模型的变化推理的语音识别方法 | |
CN1750120A (zh) | 索引设备和索引方法 | |
CN1129485A (zh) | 信号分析装置 | |
CN1298172A (zh) | 用于中等或大词汇量语音识别的上下文相关声模型 | |
CN1573926A (zh) | 用于文本和语音分类的区别性语言模型训练 | |
JPH10512686A (ja) | 個別話者に適応した音声認識のための方法及び装置 | |
CN1681002A (zh) | 语音合成系统及方法及程序产品 | |
CN1870130A (zh) | 音调模式生成方法及其装置 | |
CN1461463A (zh) | 语音合成设备 | |
CN1758263A (zh) | 基于得分差加权融合的多模态身份识别方法 | |
CN1787076A (zh) | 基于混合支持向量机的说话人识别方法 | |
CN1144172C (zh) | 包括最大似然方法的基于本征音的发言者适应方法 | |
CN1835075A (zh) | 一种结合自然样本挑选与声学参数建模的语音合成方法 | |
CN1253851C (zh) | 基于事先知识的说话者检验及说话者识别系统和方法 | |
CN1787074A (zh) | 基于情感迁移规则及语音修正的说话人识别方法 |
Legal Events
Date | Code | Title | Description |
---|---|---|---|
C06 | Publication | ||
PB01 | Publication | ||
C10 | Entry into substantive examination | ||
SE01 | Entry into force of request for substantive examination | ||
C14 | Grant of patent or utility model | ||
GR01 | Patent grant | ||
ASS | Succession or assignment of patent right |
Owner name: MATSUSHITA ELECTRIC (AMERICA) INTELLECTUAL PROPERT Free format text: FORMER OWNER: MATSUSHITA ELECTRIC INDUSTRIAL CO, LTD. Effective date: 20140714 |
|
C41 | Transfer of patent application or patent right or utility model | ||
TR01 | Transfer of patent right |
Effective date of registration: 20140714 Address after: California, USA Patentee after: PANASONIC INTELLECTUAL PROPERTY CORPORATION OF AMERICA Address before: Osaka Japan Patentee before: Matsushita Electric Industrial Co.,Ltd. |
|
CX01 | Expiry of patent term |
Granted publication date: 20041201 |
|
CX01 | Expiry of patent term |