ES2364005T3 - Procedimiento, dispositivo y medio de código de programa informático para la conversión de voz. - Google Patents

Procedimiento, dispositivo y medio de código de programa informático para la conversión de voz. Download PDF

Info

Publication number
ES2364005T3
ES2364005T3 ES08804436T ES08804436T ES2364005T3 ES 2364005 T3 ES2364005 T3 ES 2364005T3 ES 08804436 T ES08804436 T ES 08804436T ES 08804436 T ES08804436 T ES 08804436T ES 2364005 T3 ES2364005 T3 ES 2364005T3
Authority
ES
Spain
Prior art keywords
parameters
wave
vocal tract
lsf
vector
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Active
Application number
ES08804436T
Other languages
English (en)
Spanish (es)
Inventor
María Arantzazu DEL POZO ECHEZARRETA
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Del Pozo Echezarreta Maria Arantzazu
Fundacion Centro de Tecnologias de Interaccion Visual y Comunicaciones Vicomtech
Original Assignee
Del Pozo Echezarreta Maria Arantzazu
Fundacion Centro de Tecnologias de Interaccion Visual y Comunicaciones Vicomtech
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Del Pozo Echezarreta Maria Arantzazu, Fundacion Centro de Tecnologias de Interaccion Visual y Comunicaciones Vicomtech filed Critical Del Pozo Echezarreta Maria Arantzazu
Application granted granted Critical
Publication of ES2364005T3 publication Critical patent/ES2364005T3/es
Active legal-status Critical Current
Anticipated expiration legal-status Critical

Links

Classifications

    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/02Speech enhancement, e.g. noise reduction or echo cancellation
    • G10L21/0316Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude
    • G10L21/0364Speech enhancement, e.g. noise reduction or echo cancellation by changing the amplitude for improving intelligibility
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/003Changing voice quality, e.g. pitch or formants
    • G10L21/007Changing voice quality, e.g. pitch or formants characterised by the process used
    • G10L21/013Adapting to target pitch
    • G10L2021/0135Voice conversion or morphing
    • GPHYSICS
    • G10MUSICAL INSTRUMENTS; ACOUSTICS
    • G10LSPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
    • G10L21/00Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
    • G10L21/04Time compression or expansion
    • G10L21/057Time compression or expansion for improving intelligibility
    • G10L2021/0575Aids for the handicapped in speaking

Landscapes

  • Engineering & Computer Science (AREA)
  • Human Computer Interaction (AREA)
  • Acoustics & Sound (AREA)
  • Signal Processing (AREA)
  • Health & Medical Sciences (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Computational Linguistics (AREA)
  • Physics & Mathematics (AREA)
  • Quality & Reliability (AREA)
  • Multimedia (AREA)
  • Compression, Expansion, Code Conversion, And Decoders (AREA)
  • Numerical Control (AREA)
  • Auxiliary Devices For Music (AREA)
  • Circuit For Audible Band Transducer (AREA)
  • Measurement Of Mechanical Vibrations Or Ultrasonic Waves (AREA)
ES08804436T 2008-09-19 2008-09-19 Procedimiento, dispositivo y medio de código de programa informático para la conversión de voz. Active ES2364005T3 (es)

Applications Claiming Priority (1)

Application Number Priority Date Filing Date Title
PCT/EP2008/062502 WO2010031437A1 (en) 2008-09-19 2008-09-19 Method and system of voice conversion

Publications (1)

Publication Number Publication Date
ES2364005T3 true ES2364005T3 (es) 2011-08-22

Family

ID=40277465

Family Applications (1)

Application Number Title Priority Date Filing Date
ES08804436T Active ES2364005T3 (es) 2008-09-19 2008-09-19 Procedimiento, dispositivo y medio de código de programa informático para la conversión de voz.

Country Status (5)

Country Link
EP (1) EP2215632B1 (de)
AT (1) ATE502380T1 (de)
DE (1) DE602008005641D1 (de)
ES (1) ES2364005T3 (de)
WO (1) WO2010031437A1 (de)

Families Citing this family (11)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
RU2427044C1 (ru) * 2010-05-14 2011-08-20 Закрытое акционерное общество "Ай-Ти Мобайл" Текстозависимый способ конверсии голоса
CN101901598A (zh) * 2010-06-30 2010-12-01 北京捷通华声语音技术有限公司 一种哼唱合成方法和系统
ES2364401B2 (es) * 2011-06-27 2011-12-23 Universidad Politécnica de Madrid Método y sistema para la estimación de parámetros fisiológicos de la fonación.
RU2510954C2 (ru) * 2012-05-18 2014-04-10 Александр Юрьевич Бредихин Способ переозвучивания аудиоматериалов и устройство для его осуществления
US9607610B2 (en) 2014-07-03 2017-03-28 Google Inc. Devices and methods for noise modulation in a universal vocoder synthesizer
WO2020062217A1 (en) * 2018-09-30 2020-04-02 Microsoft Technology Licensing, Llc Speech waveform generation
US20220148570A1 (en) * 2019-02-25 2022-05-12 Technologies Of Voice Interface Ltd. Speech interpretation device and system
EP3839947A1 (de) 2019-12-20 2021-06-23 SoundHound, Inc. Trainieren einer stimm-morphing-vorrichtung
US11600284B2 (en) 2020-01-11 2023-03-07 Soundhound, Inc. Voice morphing apparatus having adjustable parameters
CN113780107B (zh) * 2021-08-24 2024-03-01 电信科学技术第五研究所有限公司 一种基于深度学习双输入网络模型的无线电信号检测方法
CN115641858B (zh) * 2022-10-27 2026-04-21 深圳市中科蓝讯科技股份有限公司 变声处理方法、存储介质、芯片及电子设备

Family Cites Families (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
KR100809368B1 (ko) * 2006-08-09 2008-03-05 한국과학기술원 성대파를 이용한 음색 변환 시스템

Also Published As

Publication number Publication date
DE602008005641D1 (de) 2011-04-28
WO2010031437A1 (en) 2010-03-25
ATE502380T1 (de) 2011-04-15
EP2215632A1 (de) 2010-08-11
EP2215632B1 (de) 2011-03-16

Similar Documents

Publication Publication Date Title
ES2364005T3 (es) Procedimiento, dispositivo y medio de código de programa informático para la conversión de voz.
Erro et al. Voice conversion based on weighted frequency warping
Arslan Speaker transformation algorithm using segmental codebooks (STASC)
ES2274606T3 (es) Procedimiento y aparato para obtener datos de fuente y filtro basados en formantes, para codificacion y sintesis, utilizando funcion de coste y filtrado inverso.
Childers Glottal source modeling for voice conversion
Akande et al. Estimation of the vocal tract transfer function with application to glottal wave analysis
Cabral et al. Towards an improved modeling of the glottal source in statistical parametric speech synthesis
Degottex et al. Phase minimization for glottal model estimation
Chappell et al. A comparison of spectral smoothing methods for segment concatenation based speech synthesis
Erro et al. Weighted frequency warping for voice conversion.
Cabral et al. Glottal spectral separation for parametric speech synthesis
Nercessian Differentiable WORLD synthesizer-based neural vocoder with application to end-to-end audio style transfer
Ohtsuka et al. TRANSLATED PAPER
Schleusing et al. Joint source-filter optimization for accurate vocal tract estimation using differential evolution
Degottex Glottal source and vocal-tract separation
Roebel et al. Analysis and modification of excitation source characteristics for singing voice synthesis
JP4999757B2 (ja) 音声分析合成装置、音声分析合成方法、コンピュータプログラム、および記録媒体
Lu et al. Glottal source modeling for singing voice synthesis
Agiomyrgiannakis et al. ARX-LF-based source-filter methods for voice modification and transformation
Childers et al. Factors in voice quality: acoustic features related to gender
Cabral et al. Towards a better representation of the envelope modulation of aspiration noise
Del Pozo Voice source and duration modelling for voice conversion and speech repair
Gowda et al. Quasi closed phase analysis of speech signals using time varying weighted linear prediction for accurate formant tracking
Erro et al. A pitch-asynchronous simple method for speech synthesis by diphone concatenation using the deterministic plus stochastic model
Parthasarathy et al. Articulatory analysis and synthesis of speech