TH28599A - Speech encoding equipment and methods. - Google Patents

Speech encoding equipment and methods.

Info

Publication number
TH28599A
TH28599A TH9601003594A TH9601003594A TH28599A TH 28599 A TH28599 A TH 28599A TH 9601003594 A TH9601003594 A TH 9601003594A TH 9601003594 A TH9601003594 A TH 9601003594A TH 28599 A TH28599 A TH 28599A
Authority
TH
Thailand
Prior art keywords
encoding
speech signal
signal
input speech
parameters
Prior art date
Application number
TH9601003594A
Other languages
Thai (th)
Other versions
TH28599B (en
Inventor
อีจิมา นายคาซูยูกิ
มัตสึโมโตะ นายจุน
นิชิกูจิ นายมาซายูกิ
Original Assignee
โซนี่ คอร์ปอเรชั่น
Filing date
Publication date
Application filed by โซนี่ คอร์ปอเรชั่น filed Critical โซนี่ คอร์ปอเรชั่น
Publication of TH28599A publication Critical patent/TH28599A/en
Publication of TH28599B publication Critical patent/TH28599B/en

Links

Abstract

เครื่องสำเร็จและระเบียบวิธีการเข้ารหัสเสียงพูด ซึ่งด้วยสิ่งนี้ อัตราบิตของข้อมูลที่ผ่านการเข้ารหัสแล้วสามารถได้ รับการทำให้แปรผันได้ เวกเตอร์เอาต์พุต x จะได้รับการควอน ไทซ์เวกเตอร์ที่ให้น้ำหนักโดยหน่วยการควอนไทซ์เวกเตอร์ 502 ของหน่วยการควอนไทซ์เวกเตอร์ 500 หน่วยที่หนึ่ง ดรรชนีรูป ร่างจะได้รับการเอาต์พุต ณ ขั้วต่อเอาต์พุต 503 ในขณะที่ ค่าผ่านการควอนไทซ์แล้ X0' จะได้รับการลบออกจากเวกเตอร์ของ แหล่งกำเนิด x ณ ตัวบวก 505 เวกเตอร์ของความผิดพลาดของการค วอนไทซ์ y ที่จะได้รับการแบ่งอย่างเชิงมิติโดยหน่วยการควอน ไทซ์เวกเตอร์ 510 หน่วยที่สอง ข้อมูลดรรชนีรูปร่างที่ผา นการควอนไทซ์เวกเตอร์ที่ให้น้ำหนักแล้วซึ่ง เป็นผลลัพธ์จะได้รับการเอาต์พุต ณ ขั้วต่อเอาต์ดพุต 512 1, 512 2 ค่าที่ผ่านการควอนไทซ์แล้ว Y1' และY 2' จะได้รับการ ประกอบร่วมและรวมอย่างเชิงมิติโดยตัวบวก 513 เข้ากับค่าที่ ผ่านการควอนไทซ์แล้ว X0' ค่าที่ผ่านการควอนไทซ์แล้ว X1 ที่ เป็นผลลัพธ์จะได้รับการเอาต์พุตออกมาThe finished machine and speech encoding method, with this, allows the bit rate of the encoded data to be varied. The output vector x is quantized as a weighted vector by the first quantization unit of the 502 quantization unit of the 500 quantization unit. The shape index is output at output terminal 503, while the quantized value X0' is subtracted from the source vector x at adder 505. The quantized error vector y is dimensionally divided by the second quantization unit of the 510 quantization unit. The resulting quantized shape index data, which is weighted vector, is output at output terminals 5121, 5122. The quantized values Y1' and Y2' are dimensionally combined and integrated by adder 513 with the values. After quantizing X0', the resulting quantized value X1 will be output.

Claims (9)

1. ระเบียบวิธีการเข้ารหัสเสียงพูดสำหรับการแบ่งสัญญาณเสียงพูดอินพุตบนแกนเวลาออกเป็นบล็อก ดังว่าเป็นหน่วย และ การเข้ารหัสสัญญาณที่ได้นี้ ซึ่งประกอบด้วยขั้นตอนของ : การหารีซิดิวล์ของการทำนายระยะสั้นอย่างน้อยที่สุดสำหรับ ส่วนเสียงของสัญญาณเสียงอินพุต; การหาพารามิเตอร์การเข้ารหัสชนิดการวิเคราะห์รูปไซน์ด้วย พื้นฐานบนความผิดพลาดของการทำนายระยะสั้นที่ได้รับการหา เช่นนี้;และ การดำเนินการควอนไทซ์เวกเตอร์ที่ให้น้ำหนักอย่าง เชิงสัมพัสรู้กับพารามิเตอร์การเข้ารหัสชนิดการวิเคราะห์ รูปไซน์; และการเข้ารหัสส่วนไม่มีเสียงของสัญญาณเสียง พูดอินพุตโดยการเข้ารหัสรูปคลื่น1. The speech encoding method for dividing the input speech signal on the time axis into blocks, as units, and the resulting signal encoding consists of the following steps: finding the minimum short-term prediction resistance for the annoic portion of the input speech signal; determining the sine analysis encoding parameters based on the found short-term prediction error; and performing quantization vectors that give tangential weights to the sine analysis encoding parameters; and encoding the annoic portion of the input speech signal by waveform encoding. 2. ระเบียบวิธีการเข้ารหัสสัญญาณเสียงพูดดังข้อถือสิทธิใน ข้อถือสิทธิข้อ 1 ซึ่งในที่นี้ จะตัดสินว่าสัญญาณเสียง พูดอินพุตเป็นส่วนมีเสียงหรือส่วนไม่มีเสียง และด้วยพื้น ฐานบนผลลัพธืของการตัดสิน ส่วนของสัญญาณเสียงพูดอินพุตที่ พบว่าเป็นส่วนมีเสียง จะได้รับการดำเนินกรรมวิธีด้วยการ เข้ารหัสชนิดการวิเคราะห์รูปไซน์ และส่วนของสัญญาณเสียง พูดอินพุตที่พบว่าเป็นส่วนไม่มีเสียง จะได้รับการควอนไทซ์ เวกเตอร์โดยการค้นหาเวกเตอร์ที่เหมาะที่สุดแบบลูปปิดด้วย การใช้ระเบียบวิธีการวิเคราะห์โดยการสังเคราะห์2. The speech signal encoding method described in Reputation 1 involves determining whether the input speech signal is voiced or silent. Based on this determination, the voiced portions of the input speech signal found to be voiced are processed using sine analysis encoding, while the silent portions are quantized using a closed-loop vector-optimization search with synthesis analysis. 3. ระเบียบวิธีการเข้ารหัสสัญญาณเสียงพูดดังข้อถือสิทธิใน ข้อถือสิทธิข้อ 1 ซึ่งในที่นี้ จะใช้ข้อมูลที่แสดงแทนเอ็น อีโลปเชิงสเปกตรัมเป็นพารามิเตอร์ของการวิเคราะห์รูปไซน์ ที่นำไปผ่านการควอนไทซ์เวกเตอร์ที่ให้มีน้ำหนักอย่างเชิง สัมผัสรู้ดังกล่าว3. The speech signal encoding method described in Claim 1 uses spectral representations of the elopes as parameters for the analysis of the sine wave, which is then quantized with tactile weights. 4. เครื่องสำเร็จเพื่อการเข้ารหัสเสียงพูดสำหรับการแบ่ง สัญญาณเสียงพูดอินพุตบนแกนเวลาออกเป็นบล็อก ดังว่าเป็น หน่วย และการเข้ารหัสสัญญาณที่ได้นี้ ซึ่งประกอบด้วย: อุปกรณ์สำหรับการหารีซิดิวล์ของการทำนายระยะสั้นอย่างน้อย ที่สุดของสัญญาณเสียงอินพุต; อุปกรณ์สำหรับการหาพารามิเตอร์ของการเข้ารหัสชนิดการ วิเคราะห์รูปไซน์ด้วยพื้นฐานบนรีซิดิวล์ของการทำนายระยะ สั้นที่ได้รับการหาเช่นนี้; อุปกรณ์สำหรับการดำเนินการควอนไทซ์เวกเตอร์ที่ให้น้ำหนัก อย่างเชิงสัมผัสรู้กับพารามิเตอร์การเข้ารหัสชนิดการ วิเคราะห์รูปไซด์; และ อุปกรณ์สำหรับการเข้ารหัสส่วนไม่มีเสียงของสัญญาณเสียง พูดอินพุตโดยการเข้ารหัสรูปคลื่น4. A complete speech encoding machine for dividing the input speech signal on the time axis into blocks, such as units, and encoding the resulting signal, which consists of: a device for finding the minimum short-term prediction resistance of the input speech signal; a device for determining the parameters of sine analysis encoding based on the found short-term prediction resistance; a device for performing tangent weighted vector quantization on the sine analysis encoding parameters; and a device for encoding the silent portion of the input speech signal by waveform encoding. 5.ระเบียบวิธีการเข้ารหัสสัญญาณเสียงพูดสำหรับการแบ่ง สัญญาณเสียงพูดอินพุตบนแกนเวลาให้เป็นบล็อก ดังว่าเป็น หน่วย และการเข้ารหัสสัญญาณที่ได้นี้ ซึ่งประกอบด้วยขั้น ตอนของ: การหารีซิดิวล์ของการทำนายระยะสั้นอย่างน้อยที่สุดสำหรับ ส่วนของเสียงของสัญญาณเสียงอินพุต; การหาพารามิเตอร์ของการเข้ารหัสชนิดการวิเคราะห์รูปไซน์ ด้วยพื้นฐานบนความผิดพลาดของการทำนายระยะสั้นที่ได้รับการ หาเช่นนี้; และ การดำเนินการควอนไทซ์เวกเตอร์ที่ให้น้ำหนักอย่างเชิง สัมผัสรู้กับพารามิเตอร์การเข้ารหัสชนิดการวิเคราะห์รูป ไซด์5. The speech encoding method for dividing the input speech signal on the time axis into blocks, as units, and the resulting encoding signal consists of the following steps: finding the minimum short-term prediction resistance for the audio portion of the input speech signal; determining the parameters of the sine analysis encoding based on the found short-term prediction error; and performing tactile weighting vector quantization on the sine analysis encoding parameters. 6. ระเบียบวิธีการเข้ารหัสสัญญาณเสียงพูดดังข้อถือสิทธิใน ข้อถือสิทธิข้อ 5 อย่างน้อยที่สุด จะประกอบด้วย: ขั้นตอนการควอนไทซ์เวกเตอร์ที่หนึ่ง; และ ขั้นตอนการควอนไทซ์ที่สองของการควอนไทซ์เวกเตอร์ของความ ผิดพลาดการควอนไทซ์ที่ได้รับการผลิตขึ้นมา ณ เวลาของการค วอนไทซ์เวกเตอร์ที่หนึ่งดังกล่าว6. The speech encoding procedure as provided for in Reputation 5 shall at least consist of: the first quantization vector step; and the second quantization step of the quantization vector of the quantization error produced at the time of the first quantization vector. 7. ระเบียบวิธีการเข้ารหัสสัญญาณเสียงพูดดังข้อถือสิทธิใน ข้อถือสิทธิข้อ 6 ซึ่งในนั้นสำหรับอัตราบิตที่ต่ำ เอาต์พุต ของขั้นตอนการควอนไทซ์เวกเตอร์ที่หนึ่งจะได้รับการนำออกมา และซึ่งในนั้น สำหรับอัตราบิตสูง เอาต์พุตของขั้นตอนการค วอนไทซ์เวกเตอร์ที่หนึ่งดังกล่าวและเอาต์พุตของขั้นตอนการค วอนไทซ์เวกเตอร์ที่สองดังกล่าวจะได้รับการนำออกมา7. The speech encoding method is governed by claim 6, in which, for low bit rates, the output of the first quantize vector step is taken, and in which, for high bit rates, the output of the first quantize vector step and the output of the second quantize vector step are taken. 8. เครื่องสำเร็จเพือ่การเข้ารหัสเสียงพูดสำหรับการแบ่ง สัญญาณเสียงพูดอินพุตบนแกนเวลาออกเป็นบล็อก ดังว่าเป็น หน่วย และการเข้ารหัสสัญญาณที่ได้นี้ ซึ่งประกอบด้วย: อุปกรณ์สำหรับการหารีซิดิวล์ของการทำนายระยะสั้นของสัญญาณ เสียงอินพุต; อุปกรณ์สำหรับการหาพารามิเตอร์ของการเข้ารหัสชนิดการ วิเคราะห์รูปไซน์จากความผิดพลาดของการทำนายระยะสั้นที่ได้ รับการหาเช่นนี้; อุปกรณ์สำหรับการดำเนินการควอนไทซ์เวกเตอร์แบบหลายตอนที่ ให้น้ำหนักอย่างเชิงสัมผัสรู้กับพารามิเตอร์การเข้ารหัส ชนิดการวิเคราะห์รูปไซด์8. A speech encoding device for splitting the input speech signal on the time axis into blocks, such as units, and encoding the resulting signal, which consists of: a device for finding the short-term prediction resistance of the input speech signal; a device for determining the parameters of sine analysis encoding from the obtained short-term prediction error; and a device for performing multi-stage vector quantization that tacitly weights the parameters of sine analysis encoding. 9. อุปกรณ์ปลายทางวิทยุชนิดเคลื่อนย้ายได้ ซึ่งประกอบ ด้วย; อุปกรณ์การขยายสำหรับการขยายสัญญาณเสียงพูดอินพุต; อุปกรณ์การแปลงผันเอ/ดีสำหรับการแปลงผันเอ/ดี สำหรับ สัญญาณที่ผ่านการมอดูเลตแล้วดังกล่าว; อุปกรณ์การเข้ารหัสเสียงพูดสำหรับการเข้ารหัสเอาต์พุตของ เสียงพูดของอุปกรณ์การแปลงผันเอ/ดี ดังกล่าว; อุปกรณ์การเข้ารหัสวิถีการส่งผ่านสำหรับการถอดรหัสช่อง สัญญาณที่ผ่านการเข้ารหัสแล้วที่ได้; อุปกรณ์การมอดูเลตสำหรับการมอดูเลตเอาต์พุตของอุปกรณ์การ เข้ารหัสวิถีการส่งผ่านดังกล่าว; อุปกรณ์การแปลงผันดี/เอ สำหรับการแปลงผันดี/เอ สำหรับ สัญญาณที่ผ่านการมอดูเลตแล้วที่ได้;และ อุปกรณ์ตัวขยายสำหรับการขยายสัญญาณจากอุปกรณ์การแปลงผัน ดี/เอ ดังกล่าว สำหรับการป้อนสัญญาณที่ผ่านการขยายแล้วที่ ได้ให้กับเสาอากาศ; อุปกรณ์เข้ารหัสเสียงพูดดังกล่าว ซึ่งยังประกอบต่อไปอีก ด้วย; อุปกร์สำหรับการหารีซีดิวล์ของการทำนายระยะสั้นของสัญญาณ เสียงอินพุตดังกล่าว; อุปกรณ์สำหรับการหาพารามิเตอร์ของการเข้ารหัสดชนิดการ วิเคราะห์รูปไซน์จากรีซิดิวล์ของการทำนายระยะสั้นที่ได้รับ การหาเช่นนี้; อุปกรณ์สำหรับการดำเนินการควอนไทซ์เวกเตอร์ที่ให้น้ำหนัก อย่างเชิงสัมผัสรู้กับพารามิเตอร์ของการเข้ารหัสชนิดการ วิเคราะห์รูปไซด์;และ อุปกรณ์สำหรับเข้ารหัสสัญญาเสียงพูดอินพุตดังกล่าวโดยการ เข้ารหัสรูปคลื่น (ข้อถือสิทธิ 9 ข้อ, 3 หน้า, 15 รูป)9. A portable radio terminal device comprising; an amplification device for amplifying the input speech signal; an A/D converter for A/D conversion of the modulated signal; a speech encoder for encoding the output speech of the A/D converter; a path encoder for decoding the encoded channel; a modulation device for modulating the output of the path encoder; a D/A converter for D/A conversion of the modulated signal; and an amplifier for amplifying the signal from the D/A converter for feeding the amplified signal to the antenna; the speech encoder, which further comprises; a device for determining the short-range predictive reduction of the input speech signal; a device for determining the parameters of sine analysis encoding from the determined short-range predictive reduction; a device for performing tangent weighted vector quantization on the parameters of sine analysis encoding; and a device for encoding the input speech signal by Waveform encoding (9 claims, 3 pages, 15 images)
TH9601003594A 1996-10-24 Prefixes and methods for coding speech TH28599B (en)

Publications (2)

Publication Number Publication Date
TH28599A true TH28599A (en) 1998-04-24
TH28599B TH28599B (en) 1998-04-24

Family

ID=

Similar Documents

Publication Publication Date Title
KR970022701A (en) Voice encoding method and apparatus
US6871106B1 (en) Audio signal coding apparatus, audio signal decoding apparatus, and audio signal coding and decoding apparatus
RU2233010C2 (en) Method and device for coding and decoding voice signals
JP3346765B2 (en) Audio decoding method and audio decoding device
KR970701410A (en) Sound Encoding System
JPH09127990A (en) Audio encoding method and apparatus
US5682407A (en) Voice coder for coding voice signal with code-excited linear prediction coding
JP3357829B2 (en) Audio encoding / decoding method
JPH10149199A (en) Audio encoding method, audio decoding method, audio encoding device, audio decoding device, telephone device, pitch conversion method, and medium
EP1159739A1 (en) Method and apparatus for eighth-rate random number generation for speech coders
Eriksson et al. Exploiting interframe correlation in spectral quantization: a study of different memory VQ schemes
WO2000077774A1 (en) Noise signal encoder and voice signal encoder
JP2001242896A (en) Audio encoding / decoding apparatus and method
JP3092653B2 (en) Broadband speech encoding apparatus, speech decoding apparatus, and speech encoding / decoding apparatus
JPH10240299A (en) Voice encoding and decoding device
US6678653B1 (en) Apparatus and method for coding audio data at high speed using precision information
JPH07111456A (en) Method and device for compressing voice signal
JP2796408B2 (en) Audio information compression device
JP2613503B2 (en) Speech excitation signal encoding / decoding method
WO2008118834A1 (en) Multiple stream decoder
JPH05113799A (en) Code driving linear prediction coding system
JP3496618B2 (en) Apparatus and method for speech encoding / decoding including speechless encoding operating at multiple rates
JP2898377B2 (en) Code-excited linear prediction encoder and decoder
JP3010655B2 (en) Compression encoding apparatus and method, and decoding apparatus and method
JP2762938B2 (en) Audio coding device