WO2018058425A1 - 虚拟现实引导催眠语音处理方法及装置 - Google Patents
虚拟现实引导催眠语音处理方法及装置 Download PDFInfo
- Publication number
- WO2018058425A1 WO2018058425A1 PCT/CN2016/100780 CN2016100780W WO2018058425A1 WO 2018058425 A1 WO2018058425 A1 WO 2018058425A1 CN 2016100780 W CN2016100780 W CN 2016100780W WO 2018058425 A1 WO2018058425 A1 WO 2018058425A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- hypnotic
- speech
- voice
- corpus
- guide
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Images
Classifications
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L13/00—Speech synthesis; Text to speech systems
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61M—DEVICES FOR INTRODUCING MEDIA INTO, OR ONTO, THE BODY; DEVICES FOR TRANSDUCING BODY MEDIA OR FOR TAKING MEDIA FROM THE BODY; DEVICES FOR PRODUCING OR ENDING SLEEP OR STUPOR
- A61M21/00—Other devices or methods to cause a change in the state of consciousness; Devices for producing or ending sleep by mechanical, optical, or acoustical means, e.g. for hypnosis
- A61M21/02—Other devices or methods to cause a change in the state of consciousness; Devices for producing or ending sleep by mechanical, optical, or acoustical means, e.g. for hypnosis for inducing sleep or relaxation, e.g. by direct nerve stimulation, hypnosis, analgesia
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F40/00—Handling natural language data
- G06F40/20—Natural language analysis
- G06F40/253—Grammatical analysis; Style critique
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F40/00—Handling natural language data
- G06F40/30—Semantic analysis
-
- G—PHYSICS
- G09—EDUCATION; CRYPTOGRAPHY; DISPLAY; ADVERTISING; SEALS
- G09B—EDUCATIONAL OR DEMONSTRATION APPLIANCES; APPLIANCES FOR TEACHING, OR COMMUNICATING WITH, THE BLIND, DEAF OR MUTE; MODELS; PLANETARIA; GLOBES; MAPS; DIAGRAMS
- G09B9/00—Simulators for teaching or training purposes
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L13/00—Speech synthesis; Text to speech systems
- G10L13/08—Text analysis or generation of parameters for speech synthesis out of text, e.g. grapheme to phoneme translation, prosody generation or stress or intonation determination
- G10L13/10—Prosody rules derived from text; Stress or intonation
-
- A—HUMAN NECESSITIES
- A61—MEDICAL OR VETERINARY SCIENCE; HYGIENE
- A61M—DEVICES FOR INTRODUCING MEDIA INTO, OR ONTO, THE BODY; DEVICES FOR TRANSDUCING BODY MEDIA OR FOR TAKING MEDIA FROM THE BODY; DEVICES FOR PRODUCING OR ENDING SLEEP OR STUPOR
- A61M21/00—Other devices or methods to cause a change in the state of consciousness; Devices for producing or ending sleep by mechanical, optical, or acoustical means, e.g. for hypnosis
- A61M2021/0005—Other devices or methods to cause a change in the state of consciousness; Devices for producing or ending sleep by mechanical, optical, or acoustical means, e.g. for hypnosis by the use of a particular sense, or stimulus
- A61M2021/0027—Other devices or methods to cause a change in the state of consciousness; Devices for producing or ending sleep by mechanical, optical, or acoustical means, e.g. for hypnosis by the use of a particular sense, or stimulus by the hearing sense
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L13/00—Speech synthesis; Text to speech systems
- G10L13/02—Methods for producing synthetic speech; Speech synthesisers
- G10L13/04—Details of speech synthesis systems, e.g. synthesiser structure or memory management
- G10L13/047—Architecture of speech synthesisers
Definitions
- the present invention relates to the field of virtual reality guided hypnosis technology, and in particular to a virtual reality guided hypnosis speech processing method and apparatus.
- Existing virtual reality guided hypnosis techniques are typically synthesized with a fixed, standardized hypnotic voice (recorded by a professional announcer) and a virtual reality hypnotic scene to guide the user into hypnosis.
- the user's hypnosis is guided by the fixed standardized hypnotic voice.
- it can be close to the hypnosis process voice requirements in various aspects such as tone and tone, it can not satisfy the user's faster and better hypnosis demand, and the hypnosis effect is not good.
- the embodiment of the invention provides a virtual reality guided hypnotic speech processing method for improving the hypnotic susceptibility of the user and optimizing the hypnotic effect, and the method comprises:
- the hypnotic speech data is synthesized with the virtual reality hypnotic scene, and the virtual reality is used to guide the hypnotic speech.
- the text analysis of the hypnotic guide language is performed to obtain text level information of the hypnotic guide language, including: text language, grammar and semantic analysis of the hypnotic guide language, obtaining word information, phrase information, and hypnotic guidance words, Sentence information, and relationship information between words, phrases, and sentences.
- the corpus is entered by a user susceptible according to hypnotic speech features; and/or the corpus is entered by a user susceptible at a specified sampling rate and speech resolution.
- the corpus is entered by a user susceptor using a dialect; and/or the corpus is entered by a user susceptor using a personalized language.
- the method further includes: establishing, and updating, in real time, the hypnotic speech library according to a corpus entered by a user susceptible; wherein the hidden corpus is used to disassemble the corpus in the hypnotic speech library.
- a speech unit To construct a speech unit;
- the synthesized speech unit synthesizes hypnotic speech data, including: selecting, splicing, and synthesizing the found speech units using a hidden Markov model.
- the embodiment of the invention further provides a virtual reality guided hypnosis speech processing device, which is used for improving the hypnotic susceptibility of the user and optimizing the hypnotic effect.
- the device comprises:
- a text analysis module configured to perform text analysis on the hypnotic guide language, and obtain text level information of the hypnotic guide language
- a speech analysis module configured to perform speech analysis on the hypnotic guide language, and obtain speech prosody information of the hypnotic guide language;
- a voice query module configured to search for a corresponding voice unit from the hypnotic voice library according to the text level information and the voice prosody information of the hypnotic guide language, where the hypnotic voice library stores a voice unit generated according to a corpus entered by the user susceptible person ;
- a speech synthesis module configured to synthesize hypnotic voice data by the found speech unit
- the voice output module is configured to synthesize the hypnotic voice data and the virtual reality hypnotic scene, and output the virtual reality to guide the hypnotic voice.
- the text analysis module is further configured to: perform text language, grammar, and semantic analysis on the hypnotic guide language, obtain word information, phrase information, sentence information, and words, phrases, and sentences between hypnotic guide words. Relationship information.
- the corpus is entered by a user susceptible according to hypnotic speech features; and/or the corpus is entered by a user susceptible at a specified sampling rate and speech resolution.
- the corpus is entered by a user susceptor using a dialect; and/or the corpus is entered by a user susceptor using a personalized language.
- the apparatus further includes: a corpus processing module, configured to establish, and update the hypnotic speech library in real time according to a corpus entered by a user susceptible; wherein the hypnosis is performed using a hidden Markov model
- the speech library splits the corpus to construct a speech unit
- the speech synthesis module is further configured to: select, splicing, and synthesizing the found speech units using a hidden Markov model.
- the hypnotic process speech is matched with the hypnotic susceptibility of the user, and the speech synthesis technology is used to change the original standardized hypnotic speech, and the user's sensibility speech is synthesized, and the characteristics of the hypnosis guidance are combined.
- the hypnotic voice sensitive to the user is output, thereby enhancing the hypnotic susceptibility of the user and optimizing the hypnotic effect.
- the embodiment of the invention provides a scheme for automatically synthesizing speech, which eliminates the cumbersome manual recording on the spot and satisfies the user's needs, can help the crowd without any hypnotic knowledge background to automatically output the hypnotic voice, complete the hypnosis process, and help the use. Better enter the hypnosis state.
- FIG. 1 is a schematic diagram of a virtual reality guided hypnosis voice processing method according to an embodiment of the present invention
- FIG. 2 is a schematic diagram of a specific example of a virtual reality guided hypnosis voice processing method according to an embodiment of the present invention
- FIG. 3 is a schematic diagram of a virtual reality guided hypnotic voice processing device according to an embodiment of the present invention.
- FIG. 4 is a schematic diagram of a specific example of a virtual reality guided hypnotic speech processing apparatus according to an embodiment of the present invention.
- a virtual reality guided hypnotic speech processing method which optimizes the hypnotic effect by improving the hypnotic susceptibility of the user.
- FIG. 1 is a schematic diagram of a virtual reality guided hypnosis voice processing method according to an embodiment of the present invention. As shown in FIG. 1, the method may include:
- Step 101 Perform text analysis on the hypnotic guide language to obtain text level information of the hypnosis guide language;
- Step 102 performing voice analysis on the hypnotic guide language to obtain voice prosody information of the hypnotic guide language
- Step 103 Search, according to the text level information and the phonetic prosody information of the hypnotic guide, the corresponding voice unit from the hypnotic voice library, where the hypnotic voice library stores a voice unit generated according to the corpus entered by the user susceptible;
- Step 104 Synthesize hypnotic voice data into the found voice unit
- Step 105 Synthesize the hypnotic voice data with the virtual reality hypnosis scene, and output the virtual reality to guide the hypnotic voice.
- the embodiment of the present invention fully considers that different hypnotic sound characteristics are applied to the user during the guiding hypnosis process, which will have different effects on the hypnotic effect, wherein the user is susceptible.
- the voice makes it easier to get into a specific hypnosis state for better hypnosis.
- the embodiment of the invention provides a scheme for automatically synthesizing speech, which eliminates the cumbersome manual recording on the spot, realizes the phased result that the hypnosis is completely generated by the machine, and can output the hypnotic voice with the characteristics of the user's susceptible voice. Meet user needs.
- the embodiment of the present invention can help the crowd without any hypnotic knowledge to automatically output hypnotic speech by using the speech synthesis technology, complete the hypnosis process, and help the user to enter the hypnosis state better.
- corpus gathering may be performed on the user susceptible in the early stage to establish a hypnotic speech library.
- the hypnotic voice library stores a voice unit generated based on a corpus entered by the user's susceptible person.
- the corpus can be designed according to the hypnotic speech feature to be outputted, and then the user is allowed to record the corpus under specific requirements, and then the recorded corpus is analyzed, set, and the required hypnotic speech library is established.
- the corpus entered by the user susceptible person may be entered by the user susceptible person according to the hypnotic voice feature. According to the hypnotic voice feature, the user susceptible should enter the corpus under specific requirements.
- the recorder when recording the corpus, the recorder is required to have the same volume, gentle speech, clear pronunciation, and gentle feelings.
- the corpus entered by the user's susceptible person may also be entered by the user's susceptible person at a specified sampling rate and speech resolution.
- the recorder is required to record high SNR speech at a specific sampling rate and speech resolution, making the corpus more standard.
- the embodiment of the present invention aims to solve the problem that the hypnotic effect is affected by the unfamiliarity and insensitivity of the user to the hypnotist voice in the virtual reality guiding hypnosis process by using the speech synthesis technology, and the speech synthesis technology is used to realize Automatically output the user's hypnotic sound during the virtual reality guided hypnosis process, thereby establishing an emotional connection with the user in terms of language traits, optimizing the hypnotic effect; and the user's dialect, or the voice of the person they trust is easier. It can enter a specific hypnotic state to achieve better hypnosis effect.
- the corpus entered by the user susceptor can be entered by the user's susceptor in the dialect, and/or is susceptible to the user. Entered in a personalized language. In this way, the user is satisfied by outputting a hypnotic voice with local characteristics and individuality. Demand.
- the recorder of the previous corpus suggests a specific selection.
- the hypnotic speech library is created and updated in real time according to the corpus entered by the user susceptible.
- the hidden Markov model can be used to split the corpus in the hypnotic speech library to construct a speech unit.
- the text analysis of the hypnotic guide language and the text level information of the hypnosis guide language may include, for example, performing text language, grammar and semantic analysis on the hypnotic guide language, obtaining word information, phrase information, sentence information, and hypnotic guide words. And the relationship between words, phrases, and sentences.
- the text version of the hypnotic guide can be first analyzed, and analyzed in the language layer, the grammar layer, and the semantic layer to obtain hypnosis.
- the hierarchical information of the guiding language that is, the hierarchical relationship of phrases, phrases, sentences, etc.; for example, combined with the hypnotic characteristics of virtual reality, the hypnotic guiding language completed in consultation with the professional hypnotist mainly includes progressive relaxation guidance, hypnosis scene guidance, etc.
- the text information is analyzed by grammar and semantics, and the words, phrases and sentences in the hypnotic guide are obtained.
- the prosody analysis is performed on the basis of the voice layer of the hypnosis guide, for example, analyzing the tone, tone, loudness and the like of the sound corresponding to the hypnotic guide, and obtaining prosody information at the speech level.
- the corresponding voice unit is searched from the hypnotic voice library, and then the searched voice unit is synthesized into the hypnotic voice data.
- the hidden Markov model can be used to select, splicing and synthesizing the found speech units.
- Corresponding synthesis processing is performed on the speech unit extracted from the hypnotic speech library to obtain the required speech data, that is, hypnotic speech data that is user-susceptible.
- the hypnotic leader is output in a modest and emotional manner, so in the speech synthesis process, it is necessary to control the speed and impart emotion to the hypnotic speech.
- the hypnotic speech data is synthesized with the virtual reality hypnotic scene, and the virtual reality is used to guide the hypnotic speech.
- the synthesized hypnotic voice data can be adjusted, optimized, and finally formed, and then imported into a virtual reality hypnosis scene, and the virtual reality guide hypnotic voice is output.
- FIG. 2 is a schematic diagram of a specific example of a virtual reality guided hypnosis speech processing method according to an embodiment of the present invention.
- a corpus is first designed, and the user is input by the susceptor.
- the corpus in order to establish a hypnotic speech library, using the hidden Markov model (HMM) to construct the speech unit in the hypnotic speech library
- HMM hidden Markov model
- the text analysis of the hypnotic guide, the text level information of the hypnotic guide, the speech analysis of the hypnotic guide, the speech prosody information of the hypnotic guide In the storage of the hypnotic guide, the text analysis of the hypnotic guide, the text level information of the hypnotic guide, the speech analysis of the hypnotic guide, the speech prosody information of the hypnotic guide; and then, according to the hypnosis guide Text level information and phonetic prosody information, find the corresponding phonetic unit from the hypnotic speech library, synthesize hypnotic speech data from the found speech unit; optimize the hypnotic speech data, synthesize with the virtual reality hypnosis scene, and finally output the virtual reality Guide the hypnotic voice.
- HMM hidden Markov model
- the voice synthesis technology is used in the embodiment of the present invention to collect a specific sentence of the user's hypnotic susceptible voice by collecting corpus of the user's susceptible person in the early stage, thereby establishing a hypnotic voice library, and then only It is necessary to provide text information, perform speech analysis, speech unit extraction and synthesis, and finally realize the hypnotic speech of the more sensitive person to enhance the hypnotic effect.
- the hypnosis process will increase the emotional dimension and enhance the use based on the original hypnotic effect.
- the emotional cognition of the person enhances the hypnotic effect.
- an embodiment of the present invention further provides a virtual reality guided hypnotic speech processing device, as described in the following embodiments. Since the principle of the device solving problem is similar to the virtual reality guiding hypnotic speech processing method, the implementation of the device can be referred to the implementation of the virtual reality guided hypnotic speech processing method, and the repeated description is not repeated.
- FIG. 3 is a schematic diagram of a virtual reality guided hypnotic speech processing apparatus according to an embodiment of the present invention. As shown in FIG. 3, the apparatus may include:
- a text analysis module 301 configured to perform text analysis on the hypnotic guide language, and obtain text level information of the hypnotic guide language;
- the voice analysis module 302 is configured to perform voice analysis on the hypnotic guide language to obtain voice prosody information of the hypnosis guide language;
- the voice query module 303 is configured to search for a corresponding voice unit from the hypnotic voice library according to the text layer information and the voice prosody information of the hypnosis guide word, where the hypnotic voice library stores a voice generated according to a corpus entered by the user susceptible person. unit;
- a speech synthesis module 304 configured to synthesize hypnotic voice data by the found speech unit
- the voice output module 305 is configured to synthesize the hypnotic voice data and the virtual reality hypnotic scene, and output the virtual reality to guide the hypnotic voice.
- the text analysis module 301 can be further configured to: perform text language, grammar, and semantic analysis on the hypnotic guide language, obtain word information, phrase information, sentence information, and relationship between words, phrases, and sentences in the hypnotic guide language. information.
- the corpus may be entered by the user susceptible according to the hypnotic voice feature; and/or the corpus may be entered by the user susceptor at a specified sampling rate and speech resolution.
- the corpus may be entered by a user susceptor using a dialect; and/or the corpus may be entered by a user susceptor using a personalized language.
- FIG. 4 is a schematic diagram of a specific example of a virtual reality guided hypnotic speech processing device according to an embodiment of the present invention.
- the device shown in FIG. 3 may further include: a corpus processing module 401, configured to be based on a user susceptible The entered corpus establishes and updates the hypnotic speech library in real time; wherein the hidden Markov model is used to split the corpus in the hypnotic speech library to construct a speech unit;
- the speech synthesis module 304 can be further configured to: select, splicing, and synthesizing the found speech units using a hidden Markov model.
- the speech traits of different sensitivities have different influences on the hypnotic effect of the user, and the speech synthesis technology and the virtual reality hypnosis scene are performed.
- the original standardized (recorded by a professional announcer) hypnotic voice is improved, and finally the user's hypnotic susceptible guiding voice is output, thereby achieving a more effective hypnosis state.
- the embodiment of the invention provides a scheme for automatically synthesizing speech, and can output various hypnotic voices with local characteristics to meet user requirements.
- speech synthesis technology it is possible to synthesize and output standardized hypnotic susceptibility-specific hypnotic speech, complete the hypnosis process, and help the user to enter the hypnosis state better.
- the embodiments of the present invention can be applied to a virtual reality guided hypnosis process of clinical respiratory control of radiotherapy for patients with thoracic and abdominal tumors.
- embodiments of the present invention can be provided as a method, system, or computer program product. Accordingly, the present invention may take the form of an entirely hardware embodiment, an entirely software embodiment, or a combination of software and hardware. Moreover, the invention can take the form of a computer program product embodied on one or more computer-usable storage media (including but not limited to disk storage, CD-ROM, optical storage, etc.) including computer usable program code.
- computer-usable storage media including but not limited to disk storage, CD-ROM, optical storage, etc.
- the computer program instructions can also be stored in a computer readable memory that can direct a computer or other programmable data processing device to operate in a particular manner, such that the instructions stored in the computer readable memory produce an article of manufacture comprising the instruction device.
- the apparatus implements the functions specified in one or more blocks of a flow or a flow and/or block diagram of the flowchart.
- These computer program instructions can also be loaded onto a computer or other programmable data processing device such that a series of operational steps are performed on a computer or other programmable device to produce computer-implemented processing for execution on a computer or other programmable device.
- the instructions provide steps for implementing the functions specified in one or more of the flow or in a block or blocks of a flow diagram.
Landscapes
- Engineering & Computer Science (AREA)
- Health & Medical Sciences (AREA)
- Physics & Mathematics (AREA)
- Theoretical Computer Science (AREA)
- Computational Linguistics (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Acoustics & Sound (AREA)
- General Health & Medical Sciences (AREA)
- General Physics & Mathematics (AREA)
- Anesthesiology (AREA)
- Human Computer Interaction (AREA)
- Multimedia (AREA)
- General Engineering & Computer Science (AREA)
- Artificial Intelligence (AREA)
- Pain & Pain Management (AREA)
- Psychology (AREA)
- Biomedical Technology (AREA)
- Heart & Thoracic Surgery (AREA)
- Hematology (AREA)
- Life Sciences & Earth Sciences (AREA)
- Animal Behavior & Ethology (AREA)
- Public Health (AREA)
- Veterinary Medicine (AREA)
- Educational Technology (AREA)
- Educational Administration (AREA)
- Business, Economics & Management (AREA)
- Machine Translation (AREA)
Abstract
Description
Claims (10)
- 一种虚拟现实引导催眠语音处理方法,其特征在于,包括:对催眠引导语进行文本分析,获得催眠引导语的文本层次信息;对催眠引导语进行语音分析,获得催眠引导语的语音韵律信息;根据催眠引导语的文本层次信息和语音韵律信息,从催眠语音库中查找相应的语音单元,所述催眠语音库存储有根据使用者易感者录入的语料生成的语音单元;将查找到的语音单元合成催眠语音数据;将催眠语音数据与虚拟现实催眠场景合成,输出虚拟现实引导催眠语音。
- 如权利要求1所述的方法,其特征在于,所述对催眠引导语进行文本分析,获得催眠引导语的文本层次信息,包括:对催眠引导语进行文本语言、语法及语义分析,获得催眠引导语中词语信息、词组信息、句子信息、及词语、词组、句子之间的关系信息。
- 如权利要求1所述的方法,其特征在于,所述语料是由使用者易感者根据催眠语音特征录入的;和/或,所述语料是由使用者易感者在指定的采样率和语音分辨率下录入的。
- 如权利要求1所述的方法,其特征在于,所述语料是由使用者易感者使用方言录入的;和/或,所述语料是由使用者易感者使用个性化的语言录入的。
- 如权利要求1至4任一项所述的方法,其特征在于,还包括:根据使用者易感者录入的语料建立、并实时更新所述催眠语音库;其中,使用隐马尔科夫模型在所述催眠语音库对语料进行拆分,构造语音单元;所述将查找到的语音单元合成催眠语音数据,包括:使用隐马尔科夫模型对查找到的语音单元进行挑选、拼接及合成处理。
- 一种虚拟现实引导催眠语音处理装置,其特征在于,包括:文本分析模块,用于对催眠引导语进行文本分析,获得催眠引导语的文本层次信息;语音分析模块,用于对催眠引导语进行语音分析,获得催眠引导语的语音韵律信息;语音查询模块,用于根据催眠引导语的文本层次信息和语音韵律信息,从催眠语音库中查找相应的语音单元,所述催眠语音库存储有根据使用者易感者录入的语料生成的语音单元;语音合成模块,用于将查找到的语音单元合成催眠语音数据;语音输出模块,用于将催眠语音数据与虚拟现实催眠场景合成,输出虚拟现实引导催眠语音。
- 如权利要求6所述的装置,其特征在于,所述文本分析模块进一步用于:对催眠引导语进行文本语言、语法及语义分析,获得催眠引导语中词语信息、词组信息、句子信息、及词语、词组、句子之间的关系信息。
- 如权利要求6所述的装置,其特征在于,所述语料是由使用者易感者根据催眠语音特征录入的;和/或,所述语料是由使用者易感者在指定的采样率和语音分辨率下录入的。
- 如权利要求6所述的装置,其特征在于,所述语料是由使用者易感者使用方言录入的;和/或,所述语料是由使用者易感者使用个性化的语言录入的。
- 如权利要求6至9任一项所述的装置,其特征在于,还包括:语料库处理模块,用于根据使用者易感者录入的语料建立、并实时更新所述催眠语音库;其中,使用隐马尔科夫模型在所述催眠语音库对语料进行拆分,构造语音单元;所述语音合成模块进一步用于:使用隐马尔科夫模型对查找到的语音单元进行挑选、拼接及合成处理。
Priority Applications (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| PCT/CN2016/100780 WO2018058425A1 (zh) | 2016-09-29 | 2016-09-29 | 虚拟现实引导催眠语音处理方法及装置 |
| US15/856,349 US10665221B2 (en) | 2016-09-29 | 2017-12-28 | Virtual reality guide hypnosis speech processing method and apparatus |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| PCT/CN2016/100780 WO2018058425A1 (zh) | 2016-09-29 | 2016-09-29 | 虚拟现实引导催眠语音处理方法及装置 |
Related Child Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| US15/856,349 Continuation US10665221B2 (en) | 2016-09-29 | 2017-12-28 | Virtual reality guide hypnosis speech processing method and apparatus |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2018058425A1 true WO2018058425A1 (zh) | 2018-04-05 |
Family
ID=61763557
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/CN2016/100780 Ceased WO2018058425A1 (zh) | 2016-09-29 | 2016-09-29 | 虚拟现实引导催眠语音处理方法及装置 |
Country Status (2)
| Country | Link |
|---|---|
| US (1) | US10665221B2 (zh) |
| WO (1) | WO2018058425A1 (zh) |
Families Citing this family (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20180197438A1 (en) * | 2017-01-10 | 2018-07-12 | International Business Machines Corporation | System for enhancing speech performance via pattern detection and learning |
| CN108877765A (zh) * | 2018-05-31 | 2018-11-23 | 百度在线网络技术(北京)有限公司 | 语音拼接合成的处理方法及装置、计算机设备及可读介质 |
| CN113545781B (zh) * | 2021-07-20 | 2024-06-07 | 浙江工商职业技术学院 | 虚拟现实促眠的方法及装置 |
Citations (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20010005770A1 (en) * | 1998-08-24 | 2001-06-28 | Blumenthal Richard A. | Method and apparatus to create and induce a self-created hypnosis |
| CN201286926Y (zh) * | 2008-09-26 | 2009-08-12 | 周湘峻 | 催眠仪 |
| CN102430182A (zh) * | 2011-09-01 | 2012-05-02 | 汪卫东 | 反馈式催眠治疗仪 |
| CN105536119A (zh) * | 2016-02-22 | 2016-05-04 | 南京邮电大学 | 一种基于心音控制的自适应心音催眠音乐产生方法 |
Family Cites Families (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US9649469B2 (en) * | 2008-04-24 | 2017-05-16 | The Invention Science Fund I Llc | Methods and systems for presenting a combination treatment |
-
2016
- 2016-09-29 WO PCT/CN2016/100780 patent/WO2018058425A1/zh not_active Ceased
-
2017
- 2017-12-28 US US15/856,349 patent/US10665221B2/en active Active
Patent Citations (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20010005770A1 (en) * | 1998-08-24 | 2001-06-28 | Blumenthal Richard A. | Method and apparatus to create and induce a self-created hypnosis |
| CN201286926Y (zh) * | 2008-09-26 | 2009-08-12 | 周湘峻 | 催眠仪 |
| CN102430182A (zh) * | 2011-09-01 | 2012-05-02 | 汪卫东 | 反馈式催眠治疗仪 |
| CN105536119A (zh) * | 2016-02-22 | 2016-05-04 | 南京邮电大学 | 一种基于心音控制的自适应心音催眠音乐产生方法 |
Also Published As
| Publication number | Publication date |
|---|---|
| US20180122362A1 (en) | 2018-05-03 |
| US10665221B2 (en) | 2020-05-26 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| Bigi | SPPAS-multi-lingual approaches to the automatic annotation of speech | |
| CN108806656B (zh) | 歌曲的自动生成 | |
| US8825486B2 (en) | Method and apparatus for generating synthetic speech with contrastive stress | |
| CN108806655B (zh) | 歌曲的自动生成 | |
| US8914291B2 (en) | Method and apparatus for generating synthetic speech with contrastive stress | |
| US9412359B2 (en) | System and method for cloud-based text-to-speech web services | |
| CN112382274B (zh) | 音频合成方法、装置、设备以及存储介质 | |
| Płaza et al. | Call transcription methodology for contact center systems | |
| McAuliffe et al. | ISCAN: A system for integrated phonetic analyses across speech corpora | |
| Pravena et al. | Significance of incorporating excitation source parameters for improved emotion recognition from speech and electroglottographic signals | |
| El Ouahabi et al. | Toward an automatic speech recognition system for amazigh-tarifit language | |
| Knowles et al. | Examining factors influencing the viability of automatic acoustic analysis of child speech | |
| CN119669427A (zh) | 一种双屏异显实时翻译机实现方法及设备 | |
| Kadyan et al. | Prosody features based low resource Punjabi children ASR and T-NT classifier using data augmentation | |
| CN111477210A (zh) | 语音合成方法和装置 | |
| Chen et al. | A proof-of-concept study for automatic speech recognition to transcribe AAC speakers’ speech from high-technology AAC systems | |
| CN112382269B (zh) | 音频合成方法、装置、设备以及存储介质 | |
| WO2018058425A1 (zh) | 虚拟现实引导催眠语音处理方法及装置 | |
| Labied et al. | DARIJA-C: a crowdsourced corpus for Moroccan DARIJA speech-to-text translation | |
| Kanadje et al. | Assisted keyword indexing for lecture videos using unsupervised keyword spotting | |
| CN107886938B (zh) | 虚拟现实引导催眠语音处理方法及装置 | |
| Kepuska et al. | Speech corpus generation from DVDs of movies and tv series | |
| Răgman et al. | Efficient training strategies for natural sounding speech synthesis and speaker adaptation based on FastPitch | |
| JP2012073280A (ja) | 音響モデル生成装置、音声翻訳装置、音響モデル生成方法 | |
| Long et al. | Filled pause refinement based on the pronunciation probability for lecture speech |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 16917179 Country of ref document: EP Kind code of ref document: A1 |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 32PN | Ep: public notification in the ep bulletin as address of the adressee cannot be established |
Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205A DATED 31.07.2019) |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 16917179 Country of ref document: EP Kind code of ref document: A1 |
|
| 32PN | Ep: public notification in the ep bulletin as address of the adressee cannot be established |
Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205A DATED 19-02-2020) |