TH182015A - Deciding if it is an echo / echo for speech processing. - Google Patents

Deciding if it is an echo / echo for speech processing.

Info

Publication number
TH182015A
TH182015A TH1601001123A TH1601001123A TH182015A TH 182015 A TH182015 A TH 182015A TH 1601001123 A TH1601001123 A TH 1601001123A TH 1601001123 A TH1601001123 A TH 1601001123A TH 182015 A TH182015 A TH 182015A
Authority
TH
Thailand
Prior art keywords
parameter
frame
reverberating
resonance
parameters
Prior art date
Application number
TH1601001123A
Other languages
Thai (th)
Other versions
TH1601001123A (en
Inventor
เกา
เกา นายหยาง
Original Assignee
นางดารานีย์ วัจนะวุฒิวงศ์
นางสาวสนธยา สังขพงศ์
หัวเว่ย เทคโนโลยี่ โค
Filing date
Publication date
Application filed by นางดารานีย์ วัจนะวุฒิวงศ์, นางสาวสนธยา สังขพงศ์, หัวเว่ย เทคโนโลยี่ โค filed Critical นางดารานีย์ วัจนะวุฒิวงศ์
Publication of TH182015A publication Critical patent/TH182015A/en
Publication of TH1601001123A publication Critical patent/TH1601001123A/en

Links

Abstract

ตามรูปลักษณะของการประดิษฐ์นี้ วิธีการสำหรับประมวลผลคำพูดจะรวมถึงการตัดสินกำหนด พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องที่สะท้อนลักษณะเฉพาะของคำพูดที่ไม่ก้อง/ก้องในกรอบปัจจุบัน ของสัญญาณคำพูดซึ่งประกอบรวมด้วยกรอบหลายอัน พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องที่ปรับเรียบ แล้วจะถูกตัดสินกำหนดเพื่อให้รวมถึงสารสนเทศของพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องในกรอบ ที่อยู่ก่อนหน้ากรอบปัจจุบันของสัญญาณคำพูด ผลต่างระหว่างพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องกับ พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องที่ปรับเรียบแล้วจะถูกคิดคำนวณ วิธีการยังรวมถึงการให้กำเนิด จุดตัดสินใจว่าเป็นเสียงไม่ก้อง/เสียงก้องสำหรับตัดสินกำหนดว่า กรอบปัจจุบันประกอบรวมด้วยคำพูด ที่ไม่ก้องหรือว่าคำพูดที่ก้องโดยใช้ผลต่างที่คิดคำนวณได้เป็นพารามิเตอร์ในการตัดสินใจ According to the nature of this invention The methods for processing speech include judging. The non-echo / echo parameter that reflects the characteristics of the non-echo / echo in the current frame. Of speech signals, which are made up of frames Flatten / Flatten parameters Then a decision will be determined to include information of the non-echo / echo parameters in the frame. Address preceding the current frame of speech signals. The difference between the non-echo / echo parameter with Smoothed echo / smoothed parameters will be calculated. The method also includes procreation. Decide point that it is no echo / echo for judgment. The present frame is composed of words. That do not echo, or whether words are echoed by using calculated differences as a decision parameter

Claims (17)

------24/01/2562------(OCR) หน้า 1 ของจำนวน 4 หน้า ข้อถือสิทธิ 1.วิธีการสำหรับประมวลผลคำพูด วิธีการซึ่งประกอบรวมด้วย: การตัดสินกำหนดพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนึ่งของสัญญาณคำพูดซึ่งประกอบรวมด้วยกรอบหลายอัน โดยที่พารามิเตอร์ที่ทำให้ไม่ก้องจะเป็นพารามิเตอร์แบบผสมผสานที่สะท้อนลักษณะเฉพาะอย่างน้อยสองแบบของคำพูดที่ไม่ก้องในกรอบที่หนึ่ง; การตัดสินกำหนดพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้วสำหรับกรอบที่หนึ่ง โดยที่พารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้วสำหรับกรอบที่หนึ่งนี้จะถูกคิดคำนวณจากพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้วสำหรับกรอบที่สองและพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนึ่ง โดยที่กรอบที่สองจะเป็นกรอบก่อนหน้าของกรอบที่หนึ่ง; การคิดคำนวณผลต่างระหว่างพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนึ่งกับพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้วสำหรับกรอบที่หนึ่ง; และ การตัดสินกำหนดว่ากรอบที่หนึ่งเป็นสัญญาณคำพูดที่ไม่ก้องหรือไม่โดยนำผลต่างที่คิดคำนวณได้มาเปรียบเทียบกับระดับเริ่มเปลี่ยนอย่างน้อยหนึ่งค่า 2.วิธีการของข้อถือสิทธิ 1 โดยที่ลักษณะเฉพาะอย่างน้อยสองแบบของคำพูคที่ไม่ก้องจะประกอบรวมด้วยลักษณะเฉพาะแบบมีลักษณะเป็นคาบของสัญญาณและลักษณะเฉพาะแบบการเอียงเชิงสเปกตรัม 3.วิธีการของข้อถือสิทธิ 2 โดยที่พารามิเตอร์แบบผสมผสานจะถูกคิดคำนวณจากพารามิเตอร์ลักษณะเป็นคาบและพารามิเตอร์การเอียงเชิงสเปกตรัม 4.วิธีการของข้อถือสิทธิ 1 โดยที่พารามิเตอร์ที่ทำให้ไม,ก้องที่ปรับเรียบแก้วสำหรับกรอบที่หน้งจะถูกตัดสินกำหนดโดยการให้ใ!าหนักพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนํ่งและพารามิเตอร์ที่ทำให้ไม,ก้องที'ปรับเรียบแก้วสำหรับกรอบที่สอง โดยที่ เมึ่อพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่สองมีค่ามากกว่าพารามิเตอร์ที่ทำให้1ไม่ก้องสำหรับกรอบที่หนง พารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่สองก็จะถูกให้น้ำหนักน้อยกว่าในกรณีที่พารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่สองไม่ได้มีค่ามากกว่าพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนงหน้า 2 ของจำนวน 4 หน้า5. วิธีการของข้อถือสิทธิ 4โดยที่ตัวประกอบในการใน้น้ำหนักของพารามิเตอร์ที่ทำให้ไม่กองที่ปรับเรยบแส์'วสำหรับกรอบที่สองจะเป็น 0.9 และตัวประกอบในการให้นาหนักของพารามิเตอร์ที่ทำให้ไม่น้องสำหรับกรอบที่หนั้งจะเป็น 0.1 เมื่อพารามิเตอร์ที่ทำให้ไม่น้องที่ปรับเรียบแล้วสำหรับกรอบที่สองมีค่ามากกว่าพารามิเตอร์ที่ ทำให้ไม่น้องสำหรับกรอบที่หนึ๋ง หรือโดยที่ตัวประกอบในการให้น้ำหนักของพารามิเตอร์ที่ทำให้ไม่น้องที่ปรับเรียบแล้วสำหรับกรอบที่สองจะเป็น 0.99 และตัวประกอบในการให้น้ำหนักของพารามิเตอร์ที่ทำให้ไม่น้องสำหรับกรอบที่หมื่งจะเป็น 0.01 เมื่อพารามิเตอร์ที่ทำให้ไม่น้องที่ปรับเรียบแล้วสำหรับกรอบที่สองไม่ไตัมีค่ามากกว่าพารามิเตอร์ที่ทำให้ไม่น้องสำหรับกรอบที่หนง 6. วิธีการของข้อถือสิทธิ 5 โดยที่การตัดสินกำหนดว่ากรอบที่หนึ๋งเป็นสัญญาณคำพูคที่ไม่น้องหรือไม่โดยนำผลต่างที่คิดคำนวณไค้มาเปรียบเทียบกับระดับเริ่มเปลี่ยนอย่างน้อยหนงค่านั้นจะประกอบรวมค้วย:การตัดสินกำหนดว่ากรอบที่หนงเป็นสัญญาณคำพูดที่ไม่น้องเมื่อผลต่างที่คิคคำนวณไค้มีค่ามากกว่า 0.1 หรือ การตัดสินกำหนดว่ากรอบที่หนงไม่เป็นสัญญาณคำพูดที่ไม่น้องเมื่อผลต่างที่คิคคำนวณไค้มีค่าน้อยกว่า 0.057. วิธีการของข้อถือสิทธิ 6 โดยที่วิธีการยังประกอบรวมค้วยการตัดสินกำหนดว่ากรอบที่หนี้งมีคำพูดประเภทเดียวกันกับกรอบค่อนหน้าของกรอบที่หนงเมื่อผลต่างที่คิดคำนวณไค้มีค่าไม่น้อยกว่า0.05 และไม่มากกว่า 0.1 8. วิธีการของข้อถือสิทธิ 1 โดยที่กรอบที่หนั้งและกรอบที่สองจะเป็นกรอบหรือกรอบย่อยของสัญญาณคำพูด9. ชุดเครื่องประมวลผลคำพูดชํ่งประกอบรวมค้วย:ตัวประมวลผล; และสือจัคเก็บที่อ่านไค้ค้วยคอมพิวเตอร์ที่ไม่ใช่แบบชั่วคราวซํ่งจัดเก็บคำสั่งคอมพิวเตอร์ที่เมื่อ ดำเนินการโดยตัวประมวลผลก็จะทำให้ตัวประมวลผลดำเนินการตังต่อไปนี้:หน้า 3 ของจำนวน 4 หน้าตัดสินกำหนดพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนํ่งของสัญญาณคำพูดชื่งประกอบรวมด้วยกรอบหลายอัน โดยที่พารามิเตอร์ที่ทำให้ไม,ก้องจะเป็นพารามิเตอร์แบบผสมผสานที่สะทํอนลักษณะเฉพาะอย่างน้อยสองแบบของคำพูดที่ไม่ก้องในกรอบที่หนํ่ง;ตัดสินกำหนดพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่หนง โดยที่พารามิเตอร์ ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่หนงนี้จะถูกคิดคำนวณจากพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรยบแก้วสำหรับกรอบที่สอง โดยที่กรอบที่สองจะเป็นกรอบก่อนหน้าของกรอบที่หนํ่ง;คิดคำนวณผลต่างระหว่างพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนงกับพารามิเตอร์ที่ทำให้ไม,ก้องที่ปรับเรยบแก้วสำหรับกรอบที่หนง; และตัดสินกำหนดว่ากรอบที่หนงเป็นสัญญาณคำพูคที่ไม่ก้องหรือไม่โดยนำผลต่างที่คิดคำนวณได้มา เปรยบเทียบกับระตับเรื่มเปลี่ยนอย่างน้อยหนื่งค่า10. ชุดเครื่องของก้อถือสิทธิ 9 โดยที่ลักษณะเฉพาะอย่างน้อยสองแบบของคำพูดที่ไม่ก้องจะประกอบรวมด้วยลักษณะเฉพาะแบบมีลักษณะเป็นคาบของสัญญาณและลักษณะเฉพาะแบบการเอียงเชิงสเปกตรัม11. ชุดเครื่องของก้อถือสิทธิ 10 โดยที่พารามิเตอร์แบบผสมผสานจะถูกคิดคำนวณจาก พารามิเตอร์ลักษณะเป็นคาบและพารามิเตอร์การเอียงเชิงสเปกตรัม12. ชุดเครื่องของก้อถือสิทธิ 9 ถึง 11 ข้อใดข้อหนึ๋ง โดยที่พารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่หนงจะถูกตัดสินกำหนดโดยการให้น้ำหนักพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนํ่งและพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่สอง โดยที่ เมื่อพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่สองมีค่ามากกว่าพารามิเตอร์ที่ทำให้ไม,ก้องสำหรับกรอบที่หนง พารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่สองก็จะถูกให้น้ำหนักน้อยกว่าในกรณีที่พารามิเตอร์ที่ทำให้ไม่ก้องที'ปรับเรียบแก้วสำหรับกรอบที่สองไม่ได้มีค่ามากกว่าพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนง13. ชุดเครื่องของข้อถือสิทธิ 12 โดยที่ตัวประกอบในการให้น้ำหนักของพารามิเตอร์ที่ทำให้ไม,ก้องที่ปรับเรียบแก้วสำหรับกรอบที่สองจะเป็น0.9และตัวประกอบในการให้น้ำหนักของพารามิเตอร์ ที่ทำให้ไม่ก้องสำหรับกรอบที่หนงจะเป็น 0.1 เมื่อพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแก้วสำหรับกรอบที่สองมีค่ามากกว่าพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่ห‘นง หรือหน้า4ของจำนวน4 หน'าโดยที่ตัวประกอบในการให้น้ำหนักของพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้วสำหรับกรอบที่สองจะเป็น 0.99 และตัวประกอบในการให้น้ำหนักของพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนงจะเป็น 0.01 เมื่อพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้วสำหรับกรอบที่สองไม่ได้มีค่ามากกว่าพารามิเตอร์ที่ทำให้ไม่ก้องสำหรับกรอบที่หนง 14. ชุดเครื่องของข้อถือสิทธิ 13 โดยที่กรอบที่หนึ่งจะถูกตัดสินกำหนดว่าเป็นสัญญาณคำพูดที่ไม่ก้องเมื่อผลค่างที่คิดคำนวณได้มีค่ามากกว่า 0.1 หรือกรอบที่หนึ่งจะถูกตัดสินกำหนดว่าไม่เป็นสัญญาณทำพูดที่ไม่ก้องเมื่อผลต่างที่คิดคำนวณได้มีค่าน้อยกว่า 0.0515. ชุดเครื่องของข้อถือสิทธิ 14 โดยที่ เมื่อผลต่างที่คิดคำนวณได้มีค่าระหว่าง 0.05 ถึง 0.1 กรอบที่หนึ่งก็จะลูกตัดสินกำหนดว่ามีคำพูดประเภทเคียวกันกับกรอบก่อนหน้าของกรอบที่หนึ่ง 16. ชุดเครื่องของข้อถือสิทธิ 9 ถึง 11 ข้อใดข้อหนึ่ง โดยที่กรอบที่หนึ่งและกรอบที่สองจะเป็นกรอบหรือกรอบย่อยของสัญญาณคำพูด17. อุปกรณ์เข้าถึงโสตซํ่งประกอบรวมด้วยตัวเข้ารหัสและถอดรหัสที่มีตัวเข้ารหัสหรือตัวถอดรหัส โดยที่ตัวเข้ารหัสหรือตัวถอดรหัสจะถูกจัดโครงแบบให้คำเนินวิธีการของข้อถือสิทธิ 1 ถึง 8ข้อใดข้อหนึ่ง 18. อุปกรณ์เข้าถึงโสตของข้อถือสิทธิ 17 โดยที่ตัวเข้ารหัสหรือตัวถอดรหัสจะเป็นส่วนของซิปประมวลผลสัญญาณคิจิตัล (DSP)19. ชุดเครื่องของข้อถือสิทธิ 18 โดยที่ตัวเข้ารหัสและถอดรหัสจะถูกคำเนินการโดยซอฟต์แวร์ที่คำเนินการบนตัวประมวลผล หรือโดยฮาร์ดแวร์เฉพาะงาน20. สือจัดเก็บที่อ่านได้ด้วยคอมพิวเตอร์ชํ่งจัคเก็บคำสั่งคอมพิวเตอร์ที่เมื่อคำเนินการโดย ตัวประมวลผลก็จะทำให้ตัวประมวลผลคำเนินการขั้นตอนของข้อถือสิทธิ 1 ถึง8ข้อใดข้อหนึ่ง ------------------24/01/2562------(OCR) Page 1 of 4 pages. Claims 1. Speech Processing Method. The method comprises: determination of the resonance-nosing parameter for the first frame of a speech signal composed of multiple frames, where the resonance-nosing parameter is a composite parameter reflecting at least two characteristics of the resonant speech in the first frame; determination of the smoothed resonance-nosing parameter for the first frame, where this smoothed resonance-nosing parameter for the first frame is calculated from the smoothed resonance-nosing parameter for the second frame and the resonance-nosing parameter for the first frame, where the second frame is the preceding frame of the first frame; calculation of the difference between the resonance-nosing parameter for the first frame and the smoothed resonance-nosing parameter for the first frame; and determination of whether the first frame is a resonant speech signal by comparing the calculated difference to at least one threshold level. 2. Method of Claim 1, where at least two resonance characteristics of the voice are comprised of the periodic characteristic characteristic and the spectral tilt characteristic. 3. Method of Claim 2, where the composite parameter is calculated from the periodic characteristic parameter and the spectral tilt parameter. 4. Method of Claim 1, where the glass smoothing resonance parameter for the first frame is determined by weighting the resonance parameter for the first frame and the glass smoothing resonance parameter for the second frame, where if the glass smoothing resonance parameter for the second frame is greater than the resonance parameter for the first frame, the glass smoothing resonance parameter for the second frame is given less weight than if the glass smoothing resonance parameter for the second frame is not greater than the resonance parameter for the first frame. Page 2 of 4. 5. Method of Claim. 4. Where the weighting factor of the smoothing parameter for the second frame is 0.9 and the weighting factor of the smoothing parameter for the first frame is 0.1, when the smoothing parameter for the second frame is greater than the smoothing parameter for the first frame. Or where the weighting factor of the smoothed non-smoothing parameter for the second frame is 0.99 and the weighting factor of the non-smoothing parameter for the first frame is 0.01 when the smoothed non-smoothing parameter for the second frame is not greater than the non-smoothing parameter for the first frame. 6. Method of Claim 5 where the decision of whether the first frame is a non-smoothing speech signal or not by comparing the calculated difference to at least one threshold level consists of: a decision that the first frame is a non-smoothing speech signal when the calculated difference is greater than 0.1, or a decision that the first frame is not a non-smoothing speech signal when the calculated difference is less than 0.05. 7. Method of Claim 6 where the method also consists of a decision that the first frame has the same type of speech as the preceding frame of the first frame when the calculated difference is not less than 0.05. And no more than 0.1. 8. Method of claim 1, where the first and second frames are frames or subframes of a speech signal. 9. A speech processing unit consists of: a processor; and a non-temporal computer-readable storage medium that stores computer instructions which, when executed by the processor, cause the processor to perform the following operations: Page 3 of 4. Determine the resonance-nozing parameters for the first frame of a speech signal comprising multiple frames, where the resonance-nozing parameters are composite parameters reflecting at least two characteristics of the resonance-nozing speech in the first frame; Determine the smoothing-nosonance-nozing parameters for the first frame, where these smoothing-nosonance-nozing parameters for the first frame are calculated from the smoothing-nosonance-nozing parameters for the second frame, where the second frame is the frame preceding the first frame; Calculate the difference between the resonance-nozing parameters for the first frame and the smoothing-nosonance-nozing parameters for the first frame; And determine whether the first frame is a voiceless signal or not by comparing the calculated difference to at least one initial shift level.10. The suite of rules 9, where at least two voiceless characteristics are comprised of the periodic characteristic of the signal and the spectral tilt characteristic.11. The suite of rules 10, where the composite parameter is calculated from the periodic characteristic parameter and the spectral tilt parameter.12. Any one of the suites 9 to 11, where the glass smoothing voiceless parameter for the first frame is determined by weighting the voiceless parameter for the first frame and the glass smoothing voiceless parameter for the second frame, where if the glass smoothing voiceless parameter for the second frame is greater than the voiceless parameter for the first frame, the glass smoothing voiceless parameter for the second frame is given less weight than if the glass smoothing voiceless parameter for the second frame is not greater than the voiceless parameter for the first frame.13. The set of parameters of claim 12, where the weighting factor of the reverberation-free parameter for the second frame is 0.9 and the weighting factor of the reverberation-free parameter for the first frame is 0.1, when the reverberation-free parameter for the second frame is greater than the reverberation-free parameter for the first frame, or page 4 of 4. 13. The first frame is judged to be a voiceless speech signal when the calculated difference is greater than 0.1, or the first frame is judged not to be a voiceless speech signal when the calculated difference is less than 0.05. 14. The set of claims 9 through 11 where the first and second frames are frames or subframes of the speech signal. 17. An audio access device comprising an encoder and decoder containing an encoder or decoder, where the encoder or decoder is configured to execute one of the methods of Rights 1 through 8. 18. An audio access device of Rights 17 where the encoder or decoder is part of a digital signal processing (DSP) chip. 19. A set of devices of Rights 18 where the encoder and decoder are executed by software running on a processor or by task-specific hardware. 20. A computer-readable storage medium containing computer instructions which, when executed by a processor, cause the processor to execute one of the steps of Rights 1 through 8. 1. วิธีการสำหรับประมวลผลคำพูด โดยที่วิธีการจะประกอบรวมด้วย: การตัดสินกำหนดพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องที่สะท้อนลักษณะเฉพาะของคำพูดที่ไม่ ก้อง/ก้องในกรอบปัจจุบันของสัญญาณคำพูดซึ่งประกอบรวมด้วยกรอบหลายอัน; การตัดสินกำหนดพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องที่ปรับเรียบแล้วเพื่อให้รวมถึงสารสนเทศ ของพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องในกรอบที่อยู่ก่อนหน้ากรอบปัจจุบันของสัญญาณคำพูด; การคิดคำนวณผลต่างระหว่างพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องกับพารามิเตอร์ที่ทำให้ไม่ ก้อง/ทำให้ก้องที่ปรับเรียบแล้ว; และ การตัดสินกำหนดว่า กรอบปัจจุบันประกอบรวมด้วยคำพูดที่ไม่ก้องหรือว่าคำพูดที่ก้องโดยใช้ผล ต่างที่คิดคำนวณได้เป็นพารามิเตอร์ในการตัดสินใจ1. The speech processing method comprises: determining the devouring/reverberating parameters that reflect the characteristics of the devouring/reverberating speech in the current frame of the speech signal, which consists of multiple frames; determining the smoothed devouring/reverberating parameters to include information on the devouring/reverberating parameters in the frames preceding the current frame of the speech signal; calculating the difference between the devouring/reverberating parameters and the smoothed devouring/reverberating parameters; and determining whether the current frame contains devouring or reverberating speech using the calculated difference as a decision parameter. 2. วิธีการดังระบุในข้อถือสิทธิ 1 โดยที่พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องจะเป็นพารามิเตอร์ แบบผสมผสานที่จะสะท้อนลักษณะเฉพาะอย่างน้อยสองแบบของคำพูดที่ไม่ก้อง/ก้อง2. The method described in Claim 1, where the non-reverberating/reverberating parameters are composite parameters that reflect at least two characteristics of non-reverberating/reverberating speech. 3. วิธีการดังระบุในข้อถือสิทธิ 2 โดยที่พารามิเตอร์แบบผสมผสานจะเป็นผลคูณของพารามิเตอร์ ลักษณะเป็นคาบและพารามิเตอร์การเอียงเชิงสเปกตรัม3. The method as specified in Claim 2, where the composite parameter is the product of the periodic characteristic parameter and the spectral tilt parameter. 4. วิธีการดังระบุในข้อถือสิทธิ 1 ถึง 3 ข้อใดข้อหนึ่ง โดยที่พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องจะเป็น พารามิเตอร์ที่ทำให้ไม่ก้อง (Punvoicing) ที่สะท้อนลักษณะเฉพาะของคำพูดที่ไม่ก้อง โดยที่พารามิเตอร์ที่ทำ ให้ไม่ก้อง/ทำให้ก้องที่ปรับเรียบแล้วจะเป็นพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้ว (Punvoicing_sm)4. Any of the methods specified in Claims 1 through 3, where the devoicing/reverberating parameter is a punvoicing parameter that reflects the characteristics of voiceless speech, and the smoothed devoicing/reverberating parameter is a smoothed punvoicing parameter (Punvoicing_sm). 5. วิธีการดังระบุในข้อถือสิทธิ 2 โดยที่ว่า เมื่อผลต่างระหว่างพารามิเตอร์ที่ทำให้ไม่ก้องกับพารา มิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้วจะมากกว่า 0.1 ก็ตัดสินกำหนดกรอบปัจจุบันของสัญญาณคำพูด ว่าเป็นสัญญาณที่ไม่ก้อง โดยที่ว่า เมื่อผลต่างระหว่างพารามิเตอร์ที่ทำให้ไม่ก้องกับพารามิเตอร์ที่ทำให้ไม่ ก้องที่ปรับแล้วจะน้อยกว่า 0.05 ก็ตัดสินกำหนดกรอบปัจจุบันของสัญญาณคำพูดว่าไม่เป็นคำพูดที่ไม่ ก้อง5. The method specified in Claim 2 states that when the difference between the resonance-nosing parameter and the smoothed resonance-nosing parameter is greater than 0.1, the current frame of the speech signal is deemed resonant; and when the difference between the resonance-nosing parameter and the smoothed resonance-nosing parameter is less than 0.05, the current frame of the speech signal is deemed not resonant. 6. วิธีการดังระบุในข้อถือสิทธิ 2 โดยที่ว่า เมื่อผลต่างระหว่างพารามิเตอร์ที่ทำให้ไม่ก้องกับพารา มิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้วอยู่ระหว่าง 0.05 ถึง 0.1 ก็จะตัดสินกำหนดกรอบปัจจุบันของ สัญญาณคำพูดว่ามีประเภทคำพูดที่เหมือนกันกับกรอบก่อนหน้า6. The method described in Claim 2 states that when the difference between the resonance-nosing parameter and the smoothed resonance-nosing parameter is between 0.05 and 0.1, the current frame of the speech signal is determined to have the same speech type as the previous frame. 7. วิธีการตามข้อถือสิทธิ 4 ถึง 6 ข้อใดข้อหนึ่ง โดยที่พารามิเตอร์ที่ทำให้ไม่ก้องจะถูกคิดคำนวณ จากพารามิเตอร์ที่ทำให้ไม่ก้องดังต่อไปนี้ if (Punvoicing_sm > Punvoicing) { Punvoicing_sm (สูตร) 0.9 Punvoicing_sm + 0.1 Punvoicing } else { Punvoicing_sm (สูตร) 0.99 Punvoicing_sm + 0.01 Punvoicing } .7. The method according to one of claims 4 through 6, where the resonance-free parameter is calculated from the following resonance-free parameter: if (Punvoicing_sm > Punvoicing) { Punvoicing_sm (formula) 0.9 Punvoicing_sm + 0.1 Punvoicing } else { Punvoicing_sm (formula) 0.99 Punvoicing_sm + 0.01 Punvoicing } . 8. วิธีการตามข้อถือสิทธิ 1 ถึง 3 ข้อใดข้อหนึ่ง โดยที่พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องจะเป็น พารามิเตอร์ที่ทำให้ก้อง (P voicing) ที่สะท้อนลักษณะเฉพาะของคำพูดที่ก้อง และโดยที่พารามิเตอร์ที่ทำให้ ไม่ก้อง/ทำให้ก้องที่ปรับเรียบแล้วจะเป็นพารามิเตอร์ที่ทำให้ก้องที่ปรับเรียบแล้ว (P voicing_sm)8. Any one of the claims 1 through 3 shall apply, whereby the devoicing/voicing parameter shall be the voice-reflective voice-reflective parameter (P voicing) that reflects the specific characteristics of the voiced speech, and whereby the smoothed devoicing/voicing parameter shall be the smoothed voice-reflective parameter (P voicing_sm). 9. วิธีการดังที่ระบุในข้อถือสิทธิ 8 โดยที่ว่า เมื่อผลต่างระหว่างพารามิเตอร์ที่ทำให้ก้องกับพารา มิเตอร์ที่ทำให้ก้องที่ปรับเรียบแล้วมากกว่า 0.1 ก็จะตัดสินกำหนดกรอบปัจจุบันของสัญญาณคำพูดว่าเป็น สัญญาณที่ก้อง และโดยที่ว่า เมื่อผลต่างระหว่างพารามิเตอร์ที่ทำให้ก้องกับพารามิเตอร์ที่ทำให้ก้องที่ปรับ เรียบแล้วน้อยกว่า 0.05 ก็จะตัดสินกำหนดกรอบปัจจุบันของสัญญาณตำพูดว่าไม่เป็นคำพูดที่ก้อง9. The method specified in claim 8 states that when the difference between the resonant parameter and the smoothed resonant parameter is greater than 0.1, the current frame of the speech signal is determined to be resonant; and when the difference between the resonant parameter and the smoothed resonant parameter is less than 0.05, the current frame of the speech signal is determined to not be resonant. 10. วิธีการดังระบุในข้อถือสิทธิ 8 หรือ 9 โดยที่พารามิเตอร์ที่ทำให้ก้องที่ปรับเรียบแล้วจะถูกคิด คำนวณจากพารามิเตอร์ที่ทำให้ก้องดังต่อไปนี้ if (Pvoicing_sm > Pvoicing) { Pvoicing_sm (สูตร) (7 / 8) Pvoicing_sm + (1 / 8) Pvoicing } else { Pvoicing_sm (สูตร) (255 / 256) Pvoicing_sm + (1 / 256) Pvoicing } .10. The method specified in claim 8 or 9, where the smoothed resonance parameter is calculated from the resonance parameter as follows: if (Pvoicing_sm > Pvoicing) { Pvoicing_sm (formula) (7 / 8) Pvoicing_sm + (1 / 8) Pvoicing } else { Pvoicing_sm (formula) (255 / 256) Pvoicing_sm + (1 / 256) Pvoicing } . 11. วิธีการตามข้อถือสิทธิ 1 ถึง 10 ข้อใดข้อหนึ่ง โดยที่กรอบจะประกอบรวมด้วยกรอบย่อย11. Any one of the ten claims shall be implemented in a manner that includes sub-frames. 12. ชุดเครื่องประมวลผลคำพูดซึ่งประกอบรวมด้วย: ตัวประมวลผล; และ สื่อจัดเก็บที่อ่านได้ด้วยคอมพิวเตอร์ซึ่งจัดเก็บการเขียนโปรแกรมสำหรับำเนินการโดยตัว ประมวลผลไว้ โดยที่การเขียนโปรแกรมจะรวมถึงคำสั่งให้ดำเนินการดังต่อไปนี้ : ตัดสินกำหนดพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องที่สะท้อนลักษณะเฉพาะของคำพูดที่ไม่ก้อง/ ก้องในกรอบปัจจุบันของสัญญาณคำพูดซึ่งประกอบรวมด้วยกรอบหลายอัน, ตัดสินกำหนดพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องที่ปรับเรียบแล้วเพื่อให้รวมถึงสารสนเทศของ พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องในกรอบที่อยู่ก่อนหน้ากรอบปัจจุบันของสัญญาณคำพูด, คิดคำนวณผลต่างระหว่างพารามิเตอร์ที่ทำให้ไม่ก้องทำให้ก้องกับพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำ ให้ก้องที่ปรับเรียบแล้ว, และ ตัดสินกำหนดว่า กรอบปัจจุบันประกอบรวมด้วยคำพูดที่ไม่ก้องหรือว่าคำพูดที่ก้องโดยใช้ผลต่าง ที่คิดคำนวณได้เป็นพารามิเตอร์ในการตัดสินใจ12. A speech processing unit comprising: a processor; and a computer-readable storage medium containing the programming for execution by the processor. This programming includes instructions to perform the following operations: determine the deverbinating/reverberating parameters that reflect the characteristics of the deverbinating/reverberating speech in the current frame of a speech signal composed of multiple frames; determine the smoothed deverbinating/reverberating parameters to include information on the deverbinating/reverberating parameters in frames preceding the current frame of the speech signal; calculate the difference between the deverbinating/reverberating parameters and the smoothed deverbinating/reverberating parameters; and determine whether the current frame contains deverbinating or reverberating speech using the calculated difference as a decision parameter. 13. ชุดเครื่องดังระบุในข้อถือสิทธิ 12 โดยที่พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องจะเป็นพารา มิเตอร์แบบผสมผสานที่สะท้อนผลคูณของพารามิเตอร์ลักษณะเป็นคาบและพารามิเตอร์การเอียงเชิง สเปกตรัม13. The instrument set is described in Claim 12, where the resonance/resonance-causing parameters are composite parameters reflecting the product of the periodic characteristic parameter and the spectral tilt parameter. 14. ชุดเครื่องดังระบุในข้อถือสิทธิ 12 หรือ 13 โดยที่ว่า เมื่อผลต่างระหว่างพารามิเตอร์ที่ทำให้ไม่ ก้อง/ทำให้ก้องกับพารามิเตอร์ที่ทำให้ไมาก้อง/ทำให้ก้องที่ปรับเรียบแล้วมากกว่า 0.1 ก็จะตัดสินกำหนด กรอบปัจจุบันของสัญญาณคำพูดว่า เป็นสัญญาณที่ไม่ก้อง/ก้อง โดยที่ว่า เมื่อผลต่างระหว่างพารามิเตอร์ที่ ทำให้ไม่ก้อง/ทำให้ก้องกับพารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้องที่ปรับเรียบแล้วน้อยกว่า 0.05 ก็จะตัดสิน กำหนดกรอบปัจจุบันของสัญญาณคำพูดว่าไม่เป็นคำพูดที่ไม่ก้อง/ก้อง14. The equipment specified in Claim 12 or 13 states that when the difference between the denouncing/reverberating parameter and the smoothed denouncing/reverberating parameter is greater than 0.1, the current frame of the speech signal is determined to be denouncing/reverberating; and when the difference between the denouncing/reverberating parameter and the smoothed denouncing/reverberating parameter is less than 0.05, the current frame of the speech signal is determined to be non-denouncing/reverberating. 15. ชุดเครื่องตามข้อถือสิทธิ 12 ถึง 14 ข้อใดข้อหนึ่ง โดยที่พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้อง จะเป็นพารามิเตอร์ที่ทำให้ไม่ก้องที่สะท้อนลักษณะเฉพาะของคำพูดที่ไม่ก้อง และโดยที่พารามิเตอร์ที่ทำ ให้ไม่ก้อง/ทำให้ก้องที่ปรับเรียบแล้วจะเป็นพารามิเตอร์ที่ทำให้ไม่ก้องที่ปรับเรียบแล้ว15. Any set of instruments under claims 12 through 14 where the devouring/reverberating parameters are devouring parameters that reflect the characteristics of voiceless speech, and where the smoothed devouring/reverberating parameters are smoothed devouring parameters. 16. ชุดเครื่องตามข้อถือสิทธิ 12 ถึง 14 ข้อใดข้อหนึ่ง โดยที่พารามิเตอร์ที่ทำให้ไม่ก้อง/ทำให้ก้อง จะเป็นพารามิเตอร์ที่ทำให้ก้องที่สะท้อนลักษณะเฉพาะของคำพูดที่ก้อง และโดยที่พารามิเตอร์ที่ทำให้ไม่ ก้อง/ทำให้ก้องที่ปรับเรียบแล้วจะเป็นพารามิเตอร์ที่ทำให้ก้องที่ปรับเรียบแล้ว16. Any set of instruments under claims 12 through 14 where the devouring/reverberating parameters are reverberating parameters that reflect the characteristics of reverberating speech, and where the smoothed devouring/reverberating parameters are smoothed reverberating parameters. 17. ชุดเครื่องตามข้อถือสิทธิ 13 ถึง 16 ข้อใดข้อหนึ่ง โดยที่กรอบจะประกอบรวมด้วยกรอบย่อย17. A set of instruments pursuant to any one of claims 13 through 16, where the framework comprises sub-frames.
TH1601001123A 2014-09-05 Deciding whether it is non-echoic/echoic for speech processing TH1601001123A (en)

Publications (2)

Publication Number Publication Date
TH182015A true TH182015A (en) 2018-12-13
TH1601001123A TH1601001123A (en) 2018-12-13

Family

ID=

Similar Documents

Publication Publication Date Title
JP5229234B2 (en) Non-speech segment detection method and non-speech segment detection apparatus
US12614558B2 (en) Method and apparatus for detecting correctness of pitch period
EP2843660A1 (en) Method and apparatus for detecting synthesized speech
CN105845146B (en) Method and device for voice signal processing
WO2019112468A1 (en) Multi-microphone noise reduction method, apparatus and terminal device
JP6793706B2 (en) Methods and devices for detecting audio signals
CN102667927A (en) Method and background estimator for voice activity detection
KR20170060108A (en) Neural network voice activity detection employing running range normalization
KR102012325B1 (en) Estimation of background noise in audio signals
RU2016106637A (en) DECISION ON THE AVAILABILITY / LACK OF VOCALIZATION FOR SPEECH PROCESSING
US20110196676A1 (en) Adaptive voice print for conversational biometric engine
US20150162014A1 (en) Systems and methods for enhancing an audio signal
JP7152112B2 (en) Signal processing device, signal processing method and signal processing program
CN103915099B (en) Voice fundamental periodicity detection methods and device
WO2007026436A1 (en) Vocal fry detecting device
KR20180101057A (en) Method and apparatus for voice activity detection robust to noise
US20150279373A1 (en) Voice response apparatus, method for voice processing, and recording medium having program stored thereon
JP5151103B2 (en) Voice authentication apparatus, voice authentication method and program
CN113257276A (en) Audio scene detection method, device, equipment and storage medium
Drugman et al. Detecting speech polarity with high-order statistics
US20170194018A1 (en) Noise suppression device, noise suppression method, and computer program product
Cabañas‐Molero et al. Voicing detection based on adaptive aperiodicity thresholding for speech enhancement in non‐stationary noise
KR100624439B1 (en) Voice / Unvoiced Synthesis Method
CN106373594A (en) Tone detection method and tone detection device