TH29968B - Methodologies and equipment for producing and retrieving speech signals, and methods for transmitting such speech signals. - Google Patents

Methodologies and equipment for producing and retrieving speech signals, and methods for transmitting such speech signals.

Info

Publication number
TH29968B
TH29968B TH9601002022A TH9601002022A TH29968B TH 29968 B TH29968 B TH 29968B TH 9601002022 A TH9601002022 A TH 9601002022A TH 9601002022 A TH9601002022 A TH 9601002022A TH 29968 B TH29968 B TH 29968B
Authority
TH
Thailand
Prior art keywords
speech signal
encoding parameter
encoding
input speech
sine wave
Prior art date
Application number
TH9601002022A
Other languages
Thai (th)
Other versions
TH23997A (en
Inventor
นิชิกูจิ นายมาซายูกิ
Original Assignee
โซนี่ คอร์ปอเรชั่น
นายจักรพรรดิ์ มงคลสิทธิ์
นายดำเนิน การเด่น
นายดำเนิน การเด่น นายดำเนิน การเด่น นายต่อพงศ์ โทณะวณิก นายจักรพรรดิ์ มงคลสิทธิ์ นายวิรัช ศรีเอนกราธา
นายต่อพงศ์ โทณะวณิก
นายวิรัช ศรีเอนกราธา
Filing date
Publication date
Application filed by โซนี่ คอร์ปอเรชั่น, นายจักรพรรดิ์ มงคลสิทธิ์, นายดำเนิน การเด่น, นายดำเนิน การเด่น นายดำเนิน การเด่น นายต่อพงศ์ โทณะวณิก นายจักรพรรดิ์ มงคลสิทธิ์ นายวิรัช ศรีเอนกราธา, นายต่อพงศ์ โทณะวณิก, นายวิรัช ศรีเอนกราธา filed Critical โซนี่ คอร์ปอเรชั่น
Publication of TH23997A publication Critical patent/TH23997A/en
Publication of TH29968B publication Critical patent/TH29968B/en

Links

Abstract

ชุดของรหัส จะแบ่งสัญญาณเสียงพูดซึ่งจัดเตรียมไว้ให้กับขั้วต่อ 10 ให้เป็น กรอบและจะเข้ารหัสสัญญาณที่แบ่งเอาไว้บนพื้นฐานของกรอบให้เป็นพารามิเตอร์ในการ เข้ารหัสขาออก ดังเช่นพารามิเตอร์ของคู่สเปคตรัมเชิงเส้น (แอลเอสพี) ระดับเสียงสูงต่ำ , เสียงพูด (วี) / เสียงที่ไม่ใช่เสียงพูด (ยูวี) หรือขนาดสัญญาณสเปคตรัม Am ชุดการคำนวน ค่าระหว่างกลางของพารามิเตอร์ในการเข้ารหัสที่ผ่านการดัดแปลงแล้ว 3 จะคำนวณค่า ระหว่างกลางของพารามิเตอร์ในการเข้ารหัสเพื่อที่จะคำนวณหาพารามิเตอร์ในการเข้ารหัส ที่ผ่านการดัดแปลงแล้ว ซึ่งอยู่ร่วมกับจุดของเวลาตามที่ต้องการ ชุดถอดรหัส 6 จะ สังเคราะห์คลื่นรูปไชน์ และสัญญาณรบกวน โดยมีพื้นฐานบนพารามิเตอร์ในการเข้า รหัสที่ผ่านการดัดแปลงแล้ว และจะส่งสัญญาณเสียงพูดสังเคราะห์ออกไปที่ขั้วต่อขาออก 37 การควบคุมอัตราเร็วสามารถกระทำให้เสำเร็จได้โดยง่ายมีอัตราเร็วใด ๆ ตามความ ต้องการในช่วงกว้าง พร้อมกับคุณภาพที่สูงของเสียง พร้อมด้วยปัจจัยของเสียงและ ระดับเสียงสูงต่ำที่ยังคงไม่เปลี่ยนแปลงThe code set divides the speech signal provided to connector 10 into frames and encodes the divided signal on a frame-by-frame basis into output encoding parameters, such as linear spectrum pair (LSP) parameters, pitch, speech (V)/non-speech (UV), or spectrum amplitude. The intermediate value calculation set 3 calculates the intermediate values of the encoding parameters to determine the modified encoding parameters, along with the desired time point. The decoder set 6 synthesizes a sine wave and noise based on the modified encoding parameters and outputs the synthesized speech signal to output connector 37. Speed control can be easily achieved at any desired speed over a wide range, with high audio quality, while the tone and pitch factors remain unchanged.

Claims (12)

1 ระเบียบวิธีสำหรับกาผลิตสัญญาณเสียงพูดขาเข้ากลับขึ้นมาใหม่โดยมีพื้นฐานบนพารา มิเตอร์เข้ารหัสที่หนึ่งซึ่งผลิตขึ้นโดยการแบ่งสัญญาณเสียงพูดขาเข้าออกเป็นกรอบซึ่งมีความยาวที่ กำหนดไว้ล่วงหน้าบนแกนของเวลา และโดยการเข้ารหัสสัญญาณเสียงพูดขาเข้าโดยมีพื้นฐานอยู่บน การทำทีละกรอบ , พารามิเตอร์เข้ารหัสที่หนึ่งดังกล่าวได้รับการเว้รระยะห่างโดยช่วงห่างที่หนึ่ง , ประกอบรวมด้วยขั้นตอนของ ; การผลิตพารามิเตอร์เข้ารหัสที่สอง โดยการคำนวณค่าระหว่างกลางของพารามิเตอร์เข้ารหัสที่ หนึ่งดังกล่าว , พารามิเตอร์เข้ารหัสที่สองดงกล่าวได้รับการเว้นระยะห่างโดยช่วงห่างที่สองซึ่แตกต่าง จากช่วงห่างที่หนึ่งดังกล่าว; และ การสร้างสัญญาณเสียงพูดที่ผ่านการดัดแปลง ซึ่งมีมาตราส่วนของเวลาแตกต่างจากสัญญาณ เสียงพูดขาเข้าโดยการใช้พารามิเตอร์เข้ารหัสที่สองดังกล่าว1. A method for reconstructing an input speech signal based on the first encoding parameter, which is generated by dividing the input speech signal into frames of predetermined lengths on the time axis, and by encoding the input speech signal frame by frame, the first encoding parameter is spaced by a first interval, comprising the following steps; generating a second encoding parameter by calculating the midpoint of the first encoding parameter, the second encoding parameter is spaced by a second interval different from the first interval; and constructing a modified speech signal with a different time scale from the input speech signal using the second encoding parameter. 2.ระเบียบวิธีสำหรับกาผลิตสัญญาณเสียงพูดขาเข้ากลับขึ้นมาใหม่ตามที่ถือสิทธิไว้ใน ข้อถือสิทธิที่ 1 ที่ซึ่งสัญญาณเสียงพูดที่ผ่านการดัดแปลงได้รับการผลิตขึ้นมาโดยการสังเคราะห์คลื่น รูปไซน์ โดยสอดคล้องกับพารามิเตอร์เข้ารหัสที่สองเป็นอย่างน้อยที่สุด2. A method for regenerating the input speech signal as provided in Reputation 1, in which the modified speech signal is produced by synthesizing a sine wave that satisfies at least the second encoding parameter. 3.ระเบียบวิธีสำหรับกาผลิตสัญญาณเสียงพูดขาเข้ากลับขึ้นมาใหม่ตามที่ถือสิทธิไว้ใน ข้อถือสิทธิที่ 2 ที่ซึ่งคาบเวลาของพารามิเตอร์ได้รับการเปลี่ยนแปลงโดยการดำเนินการหนึ่งของ การบีบอัดและการยืดขยายพารามิเตอร์เข้ารหัสที่หนึ่งอย่างตามลำดับ ก่อนหน้าหรือพลังขั้นตอนของ การคำนวณค่าระหว่างกลางของพารามิเตอร์เข้ารหัสที่หนึ่งดังกล่าว3. The method for regenerating the input speech signal as provided for in Claim 2, whereby the period of the parameter is changed by one operation of compression and expansion of the first encoding parameter, respectively, prior to or by the power step of the calculation of the intermediate value of such first encoding parameter. 4.ระเบียบวิธีสำหรับกาผลิตสัญญาณเสียงพูดขาเข้ากลับขึ้นมาใหม่ตามที่ถือสิทธิไว้ใน ข้อถือสิทธิที่ 1 ที่ซึ่งขั้นตอนของการคำนวณค่าระหว่างกลางของพารามิเตอร์เข้ารหัสที่หนึ่งดังกล่าว ได้รับการกระทำโดยการคำนวณค่าระหว่างกลางแบบเป็นเชิงเส้นของพารามิเตอร์ของคู่สเปคตรัม เชิงเส้น , ระดับเสียงสูงต่ำ (Pitch) และขอบเขตสเปคตรัม ตกค้างซึงมีประกอบอยู่ในพารามิเตอร์เข้า รหัสที่หนึ่งดังกล่าว4. The method for regenerating the input speech signal as provided for in Claim 1, whereby the intermediate value of the first encoding parameter is calculated by performing a linear intermediate of the parameters of the linear spectral pair, pitch, and residual spectral boundary contained in the first encoding parameter. 5.ระเบียบวิธีสำหรับกาผลิตสัญญาณเสียงพูดขาเข้ากลับขึ้นมาใหม่ตามที่ถือสิทธิไว้ใน ข้อถือสิทธิที่ 1 ที่ซึ่งพารามิเตอร์เข้ารหัสที่หนึ่งดังกล่าวที่ใช้ได้รับการกำหนดออกมาโดยการแสดงค่า ตกค้างของการทำนายระยะสั้นของสัญญาณเสียงพูดขาเข้าในลักษณะเป็นคลื่นรูปไซน์ที่ผ่านการ สังเคราะห์และสัญญาณรบกวน และโดยการเข้ารหัสสารสนเทศสเปคตรัมของความถี่ของแต่ละคลื่น รูปไซน์ที่ผ่านการสังเคราะห์และสัญญาณรบกวน5. The method for regenerating the input speech signal as provided for in Claim 1, whereby the first encoding parameter used is determined by representing the residual values of the short-term prediction of the input speech signal as a synthesized sine wave and noise, and by encoding the frequency spectrum information of each synthesized sine wave and noise. 6.เครื่องสำเร็จสำหรับกาผลิตสัญญาณเสียงพูดขาเข้ากลับขึ้นมาใหม่ ที่ซึ่งสัญญาณเสียงพูดขาเข้า ได้รับการผลิตกลับขึ้นมาใหม่โดยมีพื้นฐานบนพารามิเตอร์เข้ารหัสที่หนึ่ง ซึ่งได้รับการกำหนดออกมา โดยการแบ่งสัญญาณเสียง พูดขาเข้าออกเป็นกรอบซึ่งมีความยาวที่กำหนดไว้ล่วงหน้าบนแกนของเวลา และโดยการเข้ารหัสสัญญาณเสียงพูดขาเข้าโดยมีพื้นฐานอยู่บนการทำทีละกรอบ , พารามิเตอร์เข้า รหัสที่หนึ่งดังกล่าวได้รับการเว้นระยะห่างโดยช่วงห่างที่หนึ่ง ประกอบรวมด้วย ; อุปกรณ์การคำนวณค่าระหว่างกลางสำหรับการผลิตพารามิเตอร์เข้ารหัสที่สองโดยการ คำนวณค่าระหว่างกลางของพารามิเตอร์เข้ารหัสที่หนึ่งดังกล่าว พารามิเตอร์เข้ารหัสที่สองดังกล่าว รับการเว้นระยะห่างโดยช่วงห่างที่สองซึ่งแตกต่างจากช่วงห่างที่หนึ่งดังกล่าว และ อุปกรณ์สร้างสัญญาณเสียงพูดสำหรับการสร้างสัญญาณเสียงพูดที่ผ่านการดัดแปลง ซึ่ง มาตรส่วนของเวลาแตกต่างจากสัญญาณเสียงพูดขาเข้าโดยการใช้พารามิเตอร์เข้ารหัสที่สองดังกล่าว6. A complete device for regenerating input speech signals, where the input speech signal is regenerated based on the first encoding parameter, which is determined by dividing the input speech signal into frames of predetermined lengths on the time axis and encoding the input speech signal frame by frame. This first encoding parameter is spaced by a first interval; an intermediate value calculation device for generating the second encoding parameter by calculating the intermediate value of the first encoding parameter. This second encoding parameter is spaced by a second interval different from the first interval; and a speech signal generator for generating a modified speech signal whose time scale differs from the input speech signal using the second encoding parameter. 7.เครื่องสำเร็จการสร้างสัญญาณเสียงพูดตามที่ถือสิทธิไว้ในข้อถือสิทธิที่6ที่ซึ่งอุปกรณ์ สร้างสัญญาณเสียงพูดดังกล่าวจะสร้างสัญญษณเสียงพูดที่ผ่านการดัดแปลงดังกล่าว โดยการสังเคราะห์ คลื่นรูปไซน์ โดยสอดคล้องกับพารามิเตอร์เข้ารหัสที่สองดังกล่าวเป็นอย่างน้อยที่สุด7. A speech signal generation device as provided for in claim 6, in which such speech signal generation device will generate such modified speech signal by synthesizing a sine wave in accordance with at least the second encoding parameter. 8. เครื่องสำเร็จการสร้างสัญญาณเสียงพูดตามที่ถือสิทธิไว้ในข้อถือสิทธิที่ 7 ระกอบรวมต่อ ไปอีกด้วยอุปกรณ์การเปลี่ยนแปลงคาบเวลาที่ด้านหนึ่งของด้านต้นกระแสหรือด้านปลายกระแสของ อุปกรณ์การคำนวณค่าระหว่างกลางดังกล่าว สำหรับการบีบอัดและการยืดขยายพารามิเตอร์เข้ารหัสที่ หนึ่งดังกล่าวอย่างตามลำดับ เพื่อที่จะเปลี่ยนคาบเวลาของพารามิเตอร์เข้ารหัส8. The speech signal generation device, as entitled in claim 7, shall be further incorporated with a period-shifting device on either the upstream or downstream side of such intermediate computation device for the compression and expansion of the first encoding parameter, respectively, in order to change the period of the encoding parameter. 9. เครื่องสำเร็จการสร้างสัญญาณเสียงพูดตามที่ถือสิทธิไว้ในข้อถือสิทธิที่ 6 ที่ซึ่งอุปกรณ์ การคำนวณค่าระหว่างกลางดังกล่าวจะกระทำการคำนวณค่าระหว่างกลางแบบเป็นเชิงเส้นบนพารา มิเตอร์ของคู่สเปคตรัมเชิงเส้น ระดับเสียงสูงต่ำ และขอบเขตสเปคตรัมตกค้างซึ่งมีประกอบอยู่ใน พารามิเตอร์เข้ารหัสที่หนึ่งดังกล่าว9. A speech signal generation device as provided for in Claim 6, in which such intermediate calculation device performs linear intermediate calculations on the parameters of the linear spectral pair, pitch, and residual spectral boundary contained in the first encoding parameter. 10. เครื่องสำเร็จการสร้างสัญญาณเสียงพูดตามที่ถือสิทธิไว้ในข้อถือสิทธิที่ 6 ที่ซึ่งพารา มิเตอร์เข้ารหัสที่หนึ่งดังกล่าวที่ได้รับการกำหนดออกมาโดยการแสดงค่าตกค้างของการทำนาย ระยะสั้น ของสัญาณเสียงพูดขาเข้า ในลักษณะที่คลื่นรูปไซน์ที่ผ่านการสังเคราะห์และสัญญาณ รบกวนและโดยการเข้ารหัสข้อมูลสเปคตรัมของความถี่ของแต่ละคลื่นรูปไซน์ที่ผานการสังเคราะห์ และสัญญาณรบกวน10. A speech signal generation device as provided for in Claim 6, in which the first encoding parameter is determined by representing the residual values of the short-term prediction of the incoming speech signal in the manner of synthesized sine wave and noise, and by encoding the frequency spectrum information of each synthesized sine wave and noise. 11. ระเบียบวิธีสำหรับการส่งสัญญาณเสียงพูด ประกอบรวมด้วยขั้นตอนของ ; การผลิตพารามิเตอร์เข้ารหัสที่หนึ่งโดยการแบ่งสัญญาณเสียงพูดขาเข้าออกเป็นกรอบซึ่งมี ความยาวที่กำหนดไว้ล่วงหน้าบนแกนของเวลาและโดยการเข้ารหัสสัญญาณเสียงพูดขาเข้าโดยมี พื้นฐานอยู่บนการทำทีละกรอบ , พารามิเตอร์เข้ารหัสที่หนึ่งดังกล่าวได้รับการเว้นระยะห่างโดยช่วง ห่างที่หนึ่ง; การผลิตพารามิเตอร์เข้ารหัสที่สองโดยการคำนวณค่าระหว่างกลางของพารามิเตอร์เข้ารหัสที่ หนึ่งดังกล่าว, พารามิเตอร์เข้ารหัสที่สองดังกล่าวได้รับการเว้นระยะห่างโดยช่วงห่างที่ส่องซึ่งแตกต่าง จากช่วงห่างที่หนึ่งดังกล่าว; และ การส่งพารามิเตอร์เข้ารหัสที่สองดังกล่าว11. The methodology for transmitting speech signals consists of the following steps: ; producing the first encoding parameter by dividing the incoming speech signal into frames of a predetermined length on the time axis and by encoding the incoming speech signal on a frame-by-frame basis, such first encoding parameter is spaced by the first interval; producing the second encoding parameter by calculating the intermediate value of such first encoding parameter, such second encoding parameter is spaced by an interval different from the first interval; and transmitting such second encoding parameter. 12. ระเบียบวิธีสำหรับการส่งสัญญาณเสียงพูด ขาเข้าตามที่ถือสิทธิไว้ในข้อถือสิทธิที่ 11 ที่ ซึ่งพารามิเตอร์เข้ารหัสที่หนึ่งดังกล่าวที่ใช้ได้รับการกำหนดออกมาโดยการแสดงค่าตกค้างของการ ทำนายระยะสั้นของสัญญาณเสียงพูดขาเข้า ในลักษณะเป็นคลื่นรูปไซน์ที่ผ่านการสังเคราะห์ และ สัญญาณกบกวน และโดยการเข้ารหัสข้อมูลสเปคตรัมของความถี่ของแต่ละคลื่นรูปไซน์ที่ผ่านการ สังเคราะห์และสัญญาณรบกวน12. The method for transmitting incoming speech signals as provided for in claim 11, whereby the first encoding parameter used is determined by representing the residual values of the short-term prediction of the incoming speech signal as a synthesized sine wave and noise signal, and by encoding the frequency spectrum information of each synthesized sine wave and noise signal.
TH9601002022A 1996-06-18 Methodologies and equipment for producing and retrieving speech signals, and methods for transmitting such speech signals. TH29968B (en)

Publications (2)

Publication Number Publication Date
TH23997A TH23997A (en) 1997-03-05
TH29968B true TH29968B (en) 2011-05-06

Family

ID=

Similar Documents

Publication Publication Date Title
EP0770987B1 (en) Method and apparatus for reproducing speech signals, method and apparatus for decoding the speech, method and apparatus for synthesizing the speech and portable radio terminal apparatus
EP0673013B1 (en) Signal encoding and decoding system
EP0751493B1 (en) Method and apparatus for reproducing speech signals and method for transmitting same
US6006174A (en) Multiple impulse excitation speech encoder and decoder
RU96111955A (en) METHOD AND DEVICE FOR PLAYING SPEECH SIGNALS AND METHOD FOR THEIR TRANSMISSION
US5953697A (en) Gain estimation scheme for LPC vocoders with a shape index based on signal envelopes
WO2003010752A1 (en) Speech bandwidth extension apparatus and speech bandwidth extension method
EP0843302B1 (en) Voice coder using sinusoidal analysis and pitch control
US6023671A (en) Voiced/unvoiced decision using a plurality of sigmoid-transformed parameters for speech coding
AU669788B2 (en) Method for generating a spectral noise weighting filter for use in a speech coder
EP1385150B1 (en) Method and system for parametric characterization of transient audio signals
US5235670A (en) Multiple impulse excitation speech encoder and decoder
TH23997A (en) Methods and equipment for producing and restoring speech signals, and methods for transmitting such speech signals.
JPS6346498A (en) Prosody generation method and timing point pattern generation method
JPH01257999A (en) Voice signal encoding and decoding method, voice signal encoder and voice signal decoder
KR100310930B1 (en) Speech Synthesizer and Method Thereof
KR100421816B1 (en) A voice decoding method and a portable terminal device
EP1164577A2 (en) Method and apparatus for reproducing speech signals
JPH10232699A (en) LPC vocoder
JPH0580798A (en) Speech coding / decoding apparatus and sound source generation method
JPH034300A (en) Voice encoding and decoding system
JPS62207036A (en) Voice coding system and its apparatus
KR20120032443A (en) Method and apparatus for decoding audio signal using shaping function
JPH03132800A (en) Multi-pulse type voice encoding and decoding device
KR19980035869A (en) Speech synthesizer and method