WO2022249221A1 - 対話装置、対話方法、およびプログラム - Google Patents
対話装置、対話方法、およびプログラム Download PDFInfo
- Publication number
- WO2022249221A1 WO2022249221A1 PCT/JP2021/019515 JP2021019515W WO2022249221A1 WO 2022249221 A1 WO2022249221 A1 WO 2022249221A1 JP 2021019515 W JP2021019515 W JP 2021019515W WO 2022249221 A1 WO2022249221 A1 WO 2022249221A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- utterance
- dialogue
- user
- unit
- question
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Images
Classifications
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L13/00—Speech synthesis; Text to speech systems
- G10L13/08—Text analysis or generation of parameters for speech synthesis out of text, e.g. grapheme to phoneme translation, prosody generation or stress or intonation determination
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L13/00—Speech synthesis; Text to speech systems
- G10L13/02—Methods for producing synthetic speech; Speech synthesisers
- G10L13/027—Concept to speech synthesisers; Generation of natural phrases from machine-based concepts
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L15/00—Speech recognition
- G10L15/22—Procedures used during a speech recognition process, e.g. man-machine dialogue
Definitions
- This invention relates to technology for interacting with humans using natural language.
- Non-Patent Document 1 describes in detail the task-oriented dialogue system and the non-task-oriented dialogue system.
- Task-oriented dialogue systems are widely used as personal assistants and smart speakers on smartphones.
- the main construction methods of task-oriented dialog systems are state-transition-based and frame-based.
- a state of hearing a place name (start state), a state of hearing a date, a state of providing weather information (end state), and the like are defined.
- start state a state of hearing a place name
- end state a state of providing weather information
- the dialogue starts, it transitions to the listening state defined as the starting state.
- the listening state defined as the starting state.
- the state changes to the state of hearing the date.
- the state of listening to the date when the user speaks the date, the state transitions to the state of providing weather information.
- the weather information is conveyed to the user by referring to an external database based on the information on place names and dates heard so far, and the dialogue ends.
- This interactive act updates a "frame", an information structure internal to the system.
- the frames contain the information heard from the user from the beginning of the dialogue up to that point.
- the frame includes slots for "place name” and "date", for example. "Tomorrow” is embedded in the "Date” slot by the above dialogue action.
- Dialog control generates the next action that the dialog system should take based on the updated frame.
- actions are often expressed as dialogue acts. For example, if the "place name” slot is empty, a dialogue act having a "place name question” dialogue act type is generated.
- the system's dialogue acts are converted into natural language (eg, "Where's the weather?") by speech generation and output to the user.
- non-task-oriented dialogue systems For example, a method based on manually created response rules, an example-based method that retrieves system utterances for user utterances from large-scale texts using text retrieval methods, and response utterances based on a deep learning model based on large-scale dialogue data. There are methods to generate
- Non-Patent Documents 2 and 3 methods have been proposed for generating utterances with a consistent character by converting word endings and the like to match characters, or by referring to predetermined profile information.
- Non-Patent Document 4 a method of collecting questions and responses regarding characters from online users has been proposed (see, for example, Non-Patent Document 4). Specifically, an online user is asked to write questions about the target character, and the online user is asked to post responses to those questions. Online users have the pleasure of being able to ask questions to the characters they are interested in, and at the same time have the pleasure of imagining being able to pretend to be the characters they are interested in and responding to them.
- Non-Patent Document 4 shows that this method can efficiently collect character-like utterances from online users. It is also shown that a chat dialogue system with high character can be constructed by using pairs of collected questions and answers (hereinafter also referred to as "question-answer data").
- Non-Patent Document 4 can be used to collect a large amount of questions and their responses. Can not do it.
- a dialogue system constructed based on a small amount of question-answer data has a problem of low response capability.
- question-and-answer data is collected from online users and applied to a dialogue system, even if a large amount of data can be collected, there is a problem that exchanges beyond one question and one answer cannot be performed. For example, contextual dialogue systems that hear and respond to some information cannot be implemented.
- the object of the present invention is to use question-and-answer data to perform exchanges that exceed a single question-and-answer, and to generate highly accurate system utterances even with a small amount of question-and-answer data. to present.
- a dialogue apparatus comprises a question-and-answer collection unit that collects question-and-answer data including a dialogue state, questions, and responses, and generates an utterance template associated with the state based on the question-and-answer data.
- a template generation unit an utterance generation unit that generates system utterances using an utterance template associated with a current dialogue state, an utterance presentation unit that presents system utterances to a user, and an utterance that receives user utterances uttered by users. It includes a reception unit and a state transition unit that transitions the state of the current dialogue based on user utterances.
- a dialogue apparatus comprises a question-and-answer collection unit for collecting question-and-answer data including dialogue acts representing utterance intentions, questions, and responses;
- a template generation unit that generates a template, an utterance generation unit that generates a system utterance using an utterance template associated with the next dialogue act, an utterance presentation unit that presents the system utterance to a user, and a user who has uttered the utterance. It includes an utterance reception unit that receives utterances, and a dialogue control unit that determines the next dialogue act based on user utterances.
- a dialogue apparatus includes a question-and-answer collection unit that collects paraphrase data including an utterance and an utterance that paraphrases the utterance; A conversion model generation unit that learns an utterance conversion model that outputs utterances, an utterance generation unit that generates system utterances, and an utterance conversion unit that inputs system utterances into the utterance conversion model and obtains converted system utterances by paraphrasing the system utterances. and an utterance presentation unit that presents the post-conversion system utterance to the user.
- FIG. 1 is a diagram illustrating the functional configuration of the interactive device of the first embodiment.
- FIG. 2 is a diagram illustrating the processing procedure of the interaction method of the first embodiment.
- FIG. 3 is a diagram illustrating the functional configuration of the interactive device of the second embodiment.
- FIG. 4 is a diagram illustrating the processing procedure of the interaction method of the second embodiment.
- FIG. 5 is a diagram illustrating the functional configuration of the interactive device of the third embodiment.
- FIG. 6 is a diagram illustrating the processing procedure of the interaction method of the third embodiment.
- FIG. 7 is a diagram illustrating the functional configuration of a computer.
- the present invention collects question-response pairs associated with states and dialogue acts by asking online users to post questions and responses corresponding to states and dialogue acts, which are internal representations of the dialogue system. Then, by generating utterances based on them, the accuracy of system utterances is improved. By collecting utterances that resemble specific characters from online users, it is possible to make any dialogue system have character. In addition, by collecting character-like paraphrasing utterances from online users for responses of a given dialogue system and generating utterances based on the pair of current system utterances and character-like utterances, we can create an arbitrary dialogue system. It can have character.
- utterances are collected from online users for each state, dialogue act, and utterance, but these have different restrictions.
- the state represents the situation in which the dialogue system is placed, and there are multiple semantic contents that the dialogue system can utter in that situation.
- the utterances collected for a dialogue act are constrained by the semantic content of that dialogue act. For example, given a dialogue action of "transmitting weather information", the semantic content of utterances collected from online users must convey weather information.
- states there are cases where the semantic content is not restricted, such as "initial state of dialogue”.
- the restrictions are stricter because the base expressions are also defined. Strict restrictions mean that online users have less freedom, which leads to efficient collection of only paraphrasing necessary to realize character-likeness.
- the existing task-oriented dialogue system when a predetermined character (hereinafter referred to as "character A") is given, the existing task-oriented dialogue system is configured to respond like character A.
- a dialogue system for guiding weather information is assumed.
- Existing interactive systems for guiding weather information are state-transition-based and frame-based.
- the first embodiment is an example of a state transition-based task-oriented dialog system.
- the second and third embodiments are examples of frame-based task-oriented dialog systems.
- Each embodiment will be described with a task-oriented dialogue system as its target, but the present invention is also applicable to non-task-oriented dialogue systems as long as they have states or dialogue actions.
- character A is assumed to be a boy in elementary school. Also, a place is prepared for character A to collect questions and responses from online users. Specifically, this is a website (hereinafter referred to as a "question and answer collection site"). On the question-and-answer collection site, a user who is interested in character A can post a question about character A or a response pretending to be character A. When creating a question, tags representing states and dialogue actions can be entered as attached information.
- the first embodiment of the present invention is an example of a dialog apparatus and method for presenting system utterances for responding like character A to input user utterances in a state transition-based task-oriented dialog system.
- the dialogue device 1 of the first embodiment includes, for example, a template storage unit 10, a state extraction unit 11, a question and answer collection unit 12, a template generation unit 13, an utterance generation unit 14, an utterance presentation unit 15, A speech reception unit 16 and a state transition unit 17 are provided.
- the dialogue device 1 may include a speech recognition section 18 and a speech synthesis section 19 .
- the interaction method of the first embodiment is realized by the interaction device 1 executing the processing of each step shown in FIG.
- a dialogue device is, for example, a special device configured by reading a special program into a publicly known or dedicated computer having a central processing unit (CPU: Central Processing Unit), a main memory (RAM: Random Access Memory), etc. is.
- the interactive device executes each process under the control of, for example, a central processing unit. Data input to the interactive device and data obtained in each process are stored, for example, in a main memory device, and the data stored in the main memory device are read out to the central processing unit as necessary and used for other purposes. used for processing. At least a part of each processing unit included in the interactive device may be configured by hardware such as an integrated circuit.
- Each storage unit provided in the interactive device is, for example, a main storage device such as RAM (Random Access Memory), an auxiliary storage device composed of a hard disk, an optical disk, or a semiconductor memory device such as flash memory, or a relational database. and middleware such as a key-value store.
- a main storage device such as RAM (Random Access Memory)
- auxiliary storage device composed of a hard disk, an optical disk, or a semiconductor memory device such as flash memory, or a relational database.
- middleware such as a key-value store.
- the dialogue device 1 receives as input text representing the contents of user utterances, and outputs text representing the contents of system utterances for responding to the user utterances, thereby executing a dialogue with the user who is the dialogue partner.
- the dialogue executed by the dialogue device 1 may be text-based or speech-based.
- a dialogue screen displayed on a display unit such as a display provided in the dialogue device 1 is used to execute dialogue between the user and the dialogue device 1 .
- the display unit may be installed in the housing of the interactive device 1, or may be installed outside the housing of the interactive device 1 and connected to the interactive device 1 via a wired or wireless interface.
- the dialogue screen includes at least an input area for inputting user utterances and a display area for presenting system utterances.
- the dialogue screen may include a history area for displaying the history of the dialogue from the start of the dialogue to the present, or the history area may also serve as a display area.
- the user inputs text representing the contents of the user's utterance into the input area of the interactive screen.
- the dialogue device 1 displays text representing the content of the system utterance in the display area of the dialogue screen.
- the dialogue device 1 When executing dialogue based on speech, the dialogue device 1 further includes a speech recognition unit 18 and a speech synthesis unit 19 .
- the dialogue device 1 also has a microphone and a speaker (not shown).
- the microphone and speaker may be installed in the housing of the interactive device 1, or may be installed outside the housing of the interactive device 1 and connected to the interactive device 1 via a wired or wireless interface.
- the microphone and speaker may be installed in an android modeled after a human, or a robot modeled after an animal or a fictional character.
- an android or a robot may be provided with the speech recognition unit 18 and the speech synthesis unit 19, and the interactive device 1 may be configured to input/output text representing the contents of user utterances or system utterances.
- the microphone picks up an utterance uttered by the user and outputs a sound representing the content of the user's utterance.
- the speech recognition unit 18 receives as an input speech representing the content of user's utterance, and outputs text representing the content of user's utterance, which is the result of speech recognition of the speech.
- a text representing the content of the user's utterance is input to the utterance reception unit 16 .
- the text representing the content of the system utterance output by the utterance presentation unit 15 is input to the speech synthesis unit 19 .
- the speech synthesizing unit 19 receives a text representing the content of the system utterance, and outputs a voice representing the content of the system utterance obtained as a result of voice synthesis of the text.
- the speaker emits sound representing the content of the system utterance.
- step S11 the state extraction unit 11 acquires a list of states defined inside the dialogue device 1 (for example, the state transition unit 17) and outputs the acquired state list to the question and answer collection unit 12.
- the state transition unit 17 For example, the state transition unit 17
- step S12 the question-and-answer collection unit 12 receives the state list from the state extraction unit 11, collects question-and-answer data associated with each state from the online user, and outputs the collected question-and-answer data to the template generation unit 13. do. Specifically, first, the question-and-answer collection unit 12 adds each state as a tag to the question-and-answer collection site and makes it selectable on the posting screen. The online user selects a tag of an arbitrary state on the question-and-answer collection site, and inputs a question that character A would ask in that state and an answer to that question. As a result, the question-and-answer collection unit 12 can acquire the question-and-answer data tagged with the status.
- utterances such as "Where do you want to hear the weather?” Speech such as "When?" In the "state of providing weather information", utterances such as "### day! are collected. However, ### is a placeholder filled with weather information extracted each time from the weather information database in the utterance generation unit 14 .
- the template generator 13 receives the question-and-answer data from the question-and-answer collection unit 12, builds an utterance template from the question-and-answer data associated with each state, and stores it in the template storage unit 10.
- An utterance template is an utterance template associated with each state of the state transition model. These are used when transitioning to the relevant state. Usually, it is assumed that questions contained in question-and-answer data are used as utterance templates, but responses may be used as utterance templates. Which of the questions and answers included in the question-and-answer data is to be used as an utterance template may be determined in advance based on the content of the state.
- the utterance template for "Listening to place names” is "Where is your location?"
- the template for "Listening to dates” is "What day is it?" is "Today's weather is ###". Since an utterance template is simply a pair of a state name and an utterance, it can be constructed by selecting a state and one utterance associated with it from the collected question-answer data.
- step S14 the utterance generation unit 14 receives the current state of the dialogue as an input, acquires an utterance template associated with the current state of the dialogue from the utterance templates stored in the template storage unit 10, and obtains the acquired utterance.
- the template is used to generate text representing the content of the system utterance, and the generated text representing the content of the system utterance is output to the utterance presenting unit 15 .
- the current dialog state to be input is a predetermined start state (here, "listening to the place name”) if it is the first execution from the start of the dialog, and will be described later if it is the second or later execution. This is the state after the transition output by the state transition unit 17 .
- the information corresponding to the placeholders is obtained from a predetermined database, and by embedding the obtained information in the placeholders of the utterance template, text representing the content of the system utterance is generated. do.
- the weather information is retrieved from the weather information database (here, it is "sunny sometimes cloudy") and ### is changed to "sunny sometimes cloudy”.
- ⁇ Today's weather is sunny and sometimes cloudy'' is the text representing the content of the system utterance.
- step S15 the utterance presenting unit 15 receives the text representing the content of the system utterance from the utterance generating unit 14, and presents the text representing the content of the system utterance to the user by a predetermined method.
- the text representing the content of the system utterance is output to the display section of the dialogue device 1 .
- the dialogue is executed on a voice basis, the text representing the content of the system utterance is input to the voice synthesizing unit 18, and the voice representing the content of the system utterance outputted by the voice synthesizing unit 18 is reproduced from a predetermined speaker.
- step S100 the dialogue device 1 determines whether or not the current dialogue has ended. If it is determined that the current dialogue has not ended (NO), the process proceeds to step S16. If it is determined that the current dialogue has ended (YES), the processing is terminated and the next dialogue is waited for to start.
- the decision to end the dialogue may be made by determining whether or not the current state is a predefined end state (here, "weather information providing state").
- step S16 the speech accepting unit 16 receives the text representing the content of the user's utterance input to the dialogue device 1 (or output by the speech recognition unit 18), and uses the text representing the content of the user's utterance as a state transition signal. Output to unit 17 .
- step S17 the state transition unit 17 receives the text representing the content of the user utterance from the utterance reception unit 16, analyzes the content of the user utterance, transitions the state of the current dialogue based on the analysis result, and after the transition state is output to the utterance generation unit 14 .
- the state transition unit 17 For example, in the "listening to place name" state, if a place name is included in the user's utterance, the place name is acquired, and then the state transitions to the next "listening to date" state.
- the 'listening to the date' state if the date is included in the user's utterance, the date is obtained and then transitioned to the next 'state of providing weather information'.
- Whether or not a place name is included in the user's utterance can be determined by character string matching as to whether or not the text representing the contents of the user's utterance includes a place name that matches a list of place names prepared in advance. The same is true for dates.
- a named entity extraction technique based on a sequence labeling technique such as a conditional random field may be performed to extract place names and dates, thereby determining whether user utterances include place names and dates.
- the dialogue device 1 returns the process to step S14 and presents the system utterance associated with the post-transition state.
- the dialogue apparatus 1 repeats the presentation of system utterances (steps S14 and S15) and the acceptance of user utterances (steps S16 and S17) until it is determined in step S100 that the dialogue has ended. Run.
- a specific example of dialogue executed by the dialogue device 1 of the first embodiment is shown below.
- the first embodiment it is possible to construct a state transition-based task-oriented dialogue system for guiding weather information with a predetermined character-like utterance, as follows. Note that the description in parentheses in the system utterance represents the state at that time. System: Where do you want to hear the weather? (Listening to place names) User: I'm from Tokyo. System: When? (listening to the date) User : Today. System: It's sunny! (State of providing weather information)
- the utterance template generation unit 13 can dynamically generate an utterance template for each dialogue, so that various phrasings typical of the character A can be made. As a result, it is possible to realize a task-oriented dialog system that is more human-like, friendly, and rich in expressiveness.
- the second embodiment of the present invention is an example of a dialogue apparatus and method for presenting system utterances for responding like character A to input user utterances in a frame-based task-oriented dialogue system.
- the dialog device 2 of the second embodiment includes a template storage unit 10, a question-and-answer collection unit 12, a template generation unit 13, an utterance generation unit 14, and an utterance presentation unit provided in the dialog device 1 of the first embodiment.
- a dialogue log storage unit 20 , a dialogue action extraction unit 21 , a speech understanding unit 22 , and a dialogue control unit 23 are provided.
- the dialogue device 2 may include a speech recognition unit 18 and a speech synthesis unit 19, like the dialogue device 1 of the first embodiment.
- the interaction method of the second embodiment is realized by the interaction device 2 executing the processing of each step shown in FIG.
- the dialogue log storage unit 20 stores the dialogue log when the user and the dialogue device interacted with each other.
- the dialogue log contains text representing the contents of user utterances, text representing the contents of system utterances, and labels representing system dialogue actions.
- the system dialogue act represents the utterance intent of the system utterance and is the dialogue act type of the system dialogue act.
- the text representing the content of the user's utterance is stored when the utterance accepting unit 16 outputs the text representing the content of the user's utterance.
- the text representing the content of the system utterance and the label representing the system dialogue act are stored when the utterance generation unit 14 outputs the text representing the content of the system utterance.
- step S21 the dialogue act extraction unit 21 acquires a list of system dialogue acts from the dialogue log stored in the dialogue log storage unit 20, and outputs the acquired list of system dialogue acts to the question and answer collection unit 12.
- a list of system dialogue actions defined inside the dialogue device 2 for example, the dialogue control unit 23
- step S12 the question-and-answer collection unit 12 receives a list of system dialogue acts from the dialogue act extraction unit 21, collects question-and-answer data associated with each system dialogue act from online users, and uses the collected question-and-answer data as a template. Output to the generation unit 13 .
- the question-and-answer collection unit 12 adds each system dialogue act as a tag to the question-and-answer collection site, and makes it selectable on the posting screen.
- the online user selects an arbitrary system dialogue action tag on the question-and-answer collection site, and inputs a question that character A would ask in the system dialogue action and an answer to the question.
- the question-and-answer collection unit 12 can acquire question-and-answer data tagged with the system dialogue act. For example, utterances such as "Where do you want to hear the weather?" Utterances such as "When?" In the system dialogue act of "providing weather information", utterances such as "### day! are collected.
- the template generation unit 13 receives the question-response data from the question-response collection unit 12, constructs an utterance template from the question-response data associated with each system dialogue act, and stores it in the template storage unit 10.
- An utterance template is an utterance template associated with each system dialogue act. These are used when uttering the system dialogue act. Usually, it is assumed that questions contained in question-and-answer data are used as utterance templates, but responses may be used as utterance templates. Which of the questions and answers included in the question-and-answer data is to be used as the utterance template may be determined in advance based on the content of the dialogue act.
- the utterance template for "A question about a place name” is "Where is your place?" is "Today's weather is ###". Since an utterance template is simply a pair of a dialogue act name and an utterance, it can be constructed by selecting a system dialogue act and one associated utterance from the collected question-answer data.
- step S14 the utterance generation unit 14 receives the next system dialogue act as input, acquires an utterance template associated with the system dialogue act from the utterance templates stored in the template storage unit 10, and acquires the acquired utterance template. is used to generate a text representing the content of the system utterance, and the generated text representing the content of the system utterance is output to the utterance presentation unit 15 .
- the system dialogue act to be input is a predetermined dialogue act (for example, "question of place name”) if it is executed for the first time from the start of the dialogue, and if it is executed for the second time or later, a dialogue control unit to be described later. 23 outputs the next system interaction action.
- step S22 the utterance understanding unit 22 receives the text representing the content of the user utterance from the utterance receiving unit 16, analyzes the content of the user utterance, obtains the user dialogue act representing the intention of the user utterance, and the attribute value pair.
- the resulting user dialogue action and attribute value pair are output to the dialogue control unit 23 .
- a user interaction act is an interaction act type of a user interaction act. In the present embodiment, it is assumed that there are three user dialogue actions, namely, "transmission of location name", “transmission of date”, and “transmission of location name and date”. For example, "transmission of place names" takes place names as attributes. Propagating Date takes the date as an attribute. "Transfer place name and date” takes both place name and date as attributes.
- a user dialogue act can be obtained using a classification model learned by a machine learning method from data in which dialogue act types are assigned to utterances.
- a machine learning technique for example, logistic regression can be used, and support vector machines and neural networks can also be used.
- step S23 the dialogue control unit 23 receives the user dialogue action and the attribute value pair from the speech understanding unit 22, fills a predefined frame with the attribute value pair, and according to the state of the frame, determines the next system dialogue action to be performed. is determined, and the determined system dialogue act is output to the utterance generation unit 14 .
- the method of determining system interaction actions is performed according to rules described in the form of If-Then, for example. For example, if the user interaction action is "conveying the date", processing such as filling the "date" slot with the attribute of the date is described. Also, if there is a slot in which the value is not filled in the frame, the process of selecting the system dialogue action to inquire about that slot next is described.
- the behavior of the dialogue control part is not only the If-Then rule, but also the Encoder-Decoder type neural network that obtains the output for the input, the Markov decision process that learns the optimal action for the input, and the partially observable Markov decision. It may be implemented by reinforcement learning using processes.
- the third embodiment of the present invention is another example of a dialogue apparatus and method for presenting system utterances for responding like character A to input user utterances in a frame-based task-oriented dialogue system.
- the dialogue device 3 of the third embodiment includes a template storage unit 10, a question-and-answer collection unit 12, a template generation unit 13, an utterance generation unit 14, and an utterance presentation unit provided in the dialogue device 2 of the second embodiment.
- an utterance reception unit 16 a dialogue log storage unit 20, a dialogue act extraction unit 21, an utterance understanding unit 22, and a dialogue control unit 23, and furthermore, a transformation model storage unit 30, an utterance extraction unit 31, and a transformation model generation unit 32, and an utterance conversion unit 33.
- the dialogue device 3 may include a speech recognition unit 18 and a speech synthesis unit 19, like the dialogue device 1 of the first embodiment.
- the interaction method of the third embodiment is realized by the interaction device 3 executing the processing of each step shown in FIG.
- step S31 the utterance extraction unit 31 acquires a list of system utterances from the dialogue log stored in the dialogue log storage unit 20, and outputs the acquired list of system utterances to the question and answer collection unit 12.
- a list of system utterances that the dialogue device 3 can utter may be obtained from inside the dialogue device 3 (for example, the template storage unit 20).
- step S12-2 the question-and-answer collection unit 12 receives a list of system utterances from the utterance extraction unit 31, and pairs of each system utterance and a paraphrase utterance obtained by paraphrasing the system utterance from the online user (hereinafter, also referred to as “paraphrase data”). called), and outputs the collected paraphrasing data to the conversion model generation unit 32 .
- the question-and-answer collection unit 12 adds each system utterance to the question-and-answer collection site as a tag, and makes it selectable on the posting screen.
- the online user selects an arbitrary system utterance tag on the question-and-answer collection site, paraphrases the system utterance, and inputs the utterance character A would make.
- the question-and-answer collection unit 12 can acquire the paraphrase utterance by the character A tagged with the system utterance. For example, paraphrase utterances such as "Where do you want to hear the weather?" are collected in response to the system utterance "Where is your place?" of the system dialogue act "question about the place name.”
- the conversion model generation unit 32 receives the paraphrase data from the question-and-answer collection unit 12, and learns an utterance conversion model that paraphrases the utterance by using the tagged system utterance and the paraphrase utterance input by the online user as paired data. and stores the learned utterance conversion model in the conversion model storage unit 30 .
- a Seq2Seq model based on a neural network can be used.
- the BERT model is used for the encoder and decoder
- OpenNMT-APE is used as the tool. This tool can build a generative model that generates output utterances for inputs from tokenized pairs of utterance data.
- the speech conversion model may be learned by another method, for example, a method using a recursive neural network. BERT and OpenNMT-APE are described in detail in References 1 and 2 below.
- step S33 the utterance conversion unit 33 receives the text representing the content of the system utterance from the utterance generation unit 14, inputs the text representing the content of the system utterance to the utterance conversion model stored in the conversion model storage unit 30, A text representing the content of the system utterance after conversion is obtained by paraphrasing the system utterance, and the obtained text representing the content of the system utterance after conversion is output to the utterance presenting unit 15 .
- the utterance presenting unit 15 of the third embodiment receives the text representing the content of the converted system utterance from the utterance generating unit 14, and predetermines the text representing the content of the converted system utterance as the text representing the content of the system utterance. Present to the user in a way.
- Computer-readable recording media are, for example, non-temporary recording media such as magnetic recording devices and optical discs.
- this program will be carried out, for example, by selling, transferring, lending, etc. portable recording media such as DVDs and CD-ROMs on which the program is recorded.
- the program may be distributed by storing the program in the storage device of the server computer and transferring the program from the server computer to other computers via the network.
- a computer that executes such a program for example, first stores a program recorded on a portable recording medium or a program transferred from a server computer once in the auxiliary recording unit 1050, which is its own non-temporary storage device. Store. When executing the process, this computer reads the program stored in the auxiliary recording unit 1050, which is its own non-temporary storage device, into the storage unit 1020, which is a temporary storage device, and follows the read program. Execute the process. Also, as another execution form of this program, the computer may read the program directly from a portable recording medium and execute processing according to the program, and the program is transferred from the server computer to this computer. Each time, the processing according to the received program may be executed sequentially.
- ASP Application Service Provider
- the above-mentioned processing is executed by a so-called ASP (Application Service Provider) type service, which does not transfer the program from the server computer to this computer, and realizes the processing function only by its execution instruction and result acquisition.
- ASP Application Service Provider
- the program in this embodiment includes information that is used for processing by a computer and that conforms to the program (data that is not a direct instruction to the computer but has the property of prescribing the processing of the computer, etc.).
- the device is configured by executing a predetermined program on a computer, but at least part of these processing contents may be implemented by hardware.
Landscapes
- Engineering & Computer Science (AREA)
- Computational Linguistics (AREA)
- Health & Medical Sciences (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Human Computer Interaction (AREA)
- Physics & Mathematics (AREA)
- Acoustics & Sound (AREA)
- Multimedia (AREA)
- Information Retrieval, Db Structures And Fs Structures Therefor (AREA)
Abstract
Description
本発明では、対話システムの内部表現である状態や対話行為に対して、対応する質問と応答をオンラインユーザに投稿してもらうことで、状態や対話行為に関連付けられた質問と応答のペアを収集し、それらに基づき発話生成を行うことで、システム発話の精度を向上する。オンラインユーザから特定のキャラクタらしい発話を収集すれば、任意の対話システムにキャラクタ性を持たせることが可能となる。また、所定の対話システムの応答に対して、キャラクタらしい言い換えとなる発話をオンラインユーザから収集し、現在のシステム発話とキャラクタらしい発話のペアに基づいて発話生成を行うことで、任意の対話システムにキャラクタ性を持たせることができる。これにより、対話システムが複数の状態や対話行為を遷移するような対話を実行する場合でも、各状態や各対話行為に関連付けられた質問と応答のペアを用いることで、状況に応じて適切な応答を行うことができ、キャラクタ性を持った一問一答を超える一貫した対話を実現することができる。
この発明の第一実施形態は、状態遷移ベースのタスク指向型対話システムにおいて、入力されたユーザ発話に対して、キャラクタAらしく応答するためのシステム発話を提示する対話装置およびその方法の一例である。第一実施形態の対話装置1は、図1に示すように、例えば、テンプレート記憶部10、状態抽出部11、質問応答収集部12、テンプレート生成部13、発話生成部14、発話提示部15、発話受付部16、および状態遷移部17を備える。対話装置1は、音声認識部18および音声合成部19を備えていてもよい。この対話装置1が図2に示す各ステップの処理を実行することにより、第一実施形態の対話方法が実現される。
第一実施形態の対話装置1により実行される対話の具体例を以下に示す。第一実施形態によれば、下記のように、所定のキャラクタらしい発話で天気情報を案内するための状態遷移ベースのタスク指向型対話システムを構築することができる。なお、システム発話における括弧内の記載は、その時点での状態を表す。
システム:どこの天気が聞きたいの?(地名を聞く状態)
ユーザ :東京です。
システム:いつ?(日にちを聞く状態)
ユーザ :明日です。
システム:晴れだよ!(天気情報を提供する状態)
この発明の第二実施形態は、フレームベースのタスク指向型対話システムにおいて、入力されたユーザ発話に対して、キャラクタAらしく応答するためのシステム発話を提示する対話装置およびその方法の一例である。第二実施形態の対話装置2は、図3に示すように、第一実施形態の対話装置1が備えるテンプレート記憶部10、質問応答収集部12、テンプレート生成部13、発話生成部14、発話提示部15、および発話受付部16を備え、さらに、対話ログ記憶部20、対話行為抽出部21、発話理解部22、および対話制御部23を備える。対話装置2は、第一実施形態の対話装置1と同様に、音声認識部18および音声合成部19を備えていてもよい。この対話装置2が図4に示す各ステップの処理を実行することにより、第二実施形態の対話方法が実現される。
第二実施形態の対話装置2により実行される対話の具体例を以下に示す。第二実施形態によれば、下記のように、所定のキャラクタらしい発話で天気情報を案内するためのフレームベースのタスク指向型対話システムを構築することができる。なお、システム発話における括弧内の記載は、システム対話行為を表し、ユーザ発話における括弧内の記載は、ユーザ対話行為と属性値対を表す。※以降は対話システムの動作を説明するコメントである。
システム:どこの天気が聞きたいの?(地名の質問)※システムの初期発話として設定
ユーザ :東京です。(地名の伝達、地名=東京)
システム:いつ?(日付の質問)
ユーザ :明日です。(日付の伝達、日付=明日)
システム:晴れだよ!(天気情報の提供)
この発明の第三実施形態は、フレームベースのタスク指向型対話システムにおいて、入力されたユーザ発話に対して、キャラクタAらしく応答するためのシステム発話を提示する対話装置およびその方法の他の例である。第三実施形態の対話装置3は、図5に示すように、第二実施形態の対話装置2が備えるテンプレート記憶部10、質問応答収集部12、テンプレート生成部13、発話生成部14、発話提示部15、発話受付部16、対話ログ記憶部20、対話行為抽出部21、発話理解部22、および対話制御部23を備え、さらに、変換モデル記憶部30、発話抽出部31、変換モデル生成部32、および発話変換部33を備える。対話装置3は、第一実施形態の対話装置1と同様に、音声認識部18および音声合成部19を備えていてもよい。この対話装置3が図6に示す各ステップの処理を実行することにより、第三実施形態の対話方法が実現される。
〔参考文献2〕Gon,calo M. Correia, Andre F. T. Martins, "A Simple and Effective Approach to Automatic Post-Editing with Transfer Learning," Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, 2019.
第三実施形態の対話装置3により実行される対話の具体例を以下に示す。第三実施形態によれば、下記のように、所定のキャラクタらしい発話で天気情報を案内するためのフレームベースのタスク指向型対話システムを構築することができる。なお、システム発話における括弧内の記載は、システム対話行為を表し、ユーザ発話における括弧内の記載は、ユーザ対話行為と属性値対を表す。※以降は対話システムの動作を説明するコメントである。
システム:どこの天気が聞きたいの?(地名の質問)※システムの初期発話として設定
ユーザ :東京です。(地名の伝達、地名=東京)
システム:いつ?(日付の質問)※「いつですか?」を「いつ?」に言い換え
ユーザ :明日です。(日付の伝達、日付=明日)
システム:晴れだよ!(天気情報の提供)※「晴れです」を「晴れだよ!」に言い換え
本発明により、オンラインユーザから収集できた質問応答データが少なかったとしても、対話システムの内部表現である状態や対話行為に基づいてシステム発話を生成するため、対話の状況に応じて適切なシステム発話を提示することができる。オンラインユーザから特定のキャラクタらしい発話を収集すれば、既存の対話システムにキャラクタ性を持たせることができるようになり、システム開発者が対象となるキャラクタ向けに発話生成部を作り直す必要がなくなる。加えて、対話システムの状態や対話行為に紐づいた質問応答データを収集し、予め対話システムが有する状態や対話行為の遷移と組み合わせることにより、一問一答を超え、かつ、キャラクタらしいやり取りが可能となる。
上記実施形態で説明した各装置における各種の処理機能をコンピュータによって実現する場合、各装置が有すべき機能の処理内容はプログラムによって記述される。そして、このプログラムを図7に示すコンピュータの記憶部1020に読み込ませ、演算処理部1010、入力部1030、出力部1040などに動作させることにより、上記各装置における各種の処理機能がコンピュータ上で実現される。
Claims (8)
- 対話の状態と質問と応答とを含む質問応答データを収集する質問応答収集部と、
前記質問応答データに基づいて前記状態と関連付けられた発話テンプレートを生成するテンプレート生成部と、
現在の対話の状態に関連付けられた前記発話テンプレートを用いてシステム発話を生成する発話生成部と、
前記システム発話をユーザへ提示する発話提示部と、
前記ユーザが発話したユーザ発話を受け付ける発話受付部と、
前記ユーザ発話に基づいて前記現在の対話の状態を遷移させる状態遷移部と、
を含む対話装置。 - 発話意図を表す対話行為と質問と応答とを含む質問応答データを収集する質問応答収集部と、
前記質問応答データに基づいて前記対話行為と関連付けられた発話テンプレートを生成するテンプレート生成部と、
次に行う対話行為に関連付けられた前記発話テンプレートを用いてシステム発話を生成する発話生成部と、
前記システム発話をユーザへ提示する発話提示部と、
前記ユーザが発話したユーザ発話を受け付ける発話受付部と、
前記ユーザ発話に基づいて前記次に行う対話行為を決定する対話制御部と、
を含む対話装置。 - 請求項2に記載の対話装置であって、
前記システム発話とそのシステム発話を言い換えた発話とを含む言い替えデータを用いて、発話を入力とし、その発話を言い換えた発話を出力する発話変換モデルを学習する変換モデル生成部と、
前記システム発話を前記発話変換モデルに入力して前記システム発話を言い換えた変換後システム発話を得る発話変換部と、
をさらに含む対話装置。 - 発話とその発話を言い換えた発話とを含む言い替えデータを収集する質問応答収集部と、
前記言い替えデータを用いて、発話を入力とし、その発話を言い換えた発話を出力する発話変換モデルを学習する変換モデル生成部と、
システム発話を生成する発話生成部と、
前記システム発話を前記発話変換モデルに入力して前記システム発話を言い換えた変換後システム発話を得る発話変換部と、
前記変換後システム発話をユーザへ提示する発話提示部と、
を含む対話装置。 - 質問応答収集部が、対話の状態と質問と応答とを含む質問応答データを収集し、
テンプレート生成部が、前記質問応答データに基づいて前記状態と関連付けられた発話テンプレートを生成し、
発話生成部が、現在の対話の状態に関連付けられた前記発話テンプレートを用いてシステム発話を生成し、
発話提示部が、前記システム発話をユーザへ提示し、
発話受付部が、前記ユーザが発話したユーザ発話を受け付け、
状態遷移部が、前記ユーザ発話に基づいて前記現在の対話の状態を遷移させる、
対話方法。 - 質問応答収集部が、発話意図を表す対話行為と質問と応答とを含む質問応答データを収集し、
テンプレート生成部が、前記質問応答データに基づいて前記対話行為と関連付けられた発話テンプレートを生成し、
発話生成部が、次に行う対話行為に関連付けられた前記発話テンプレートを用いてシステム発話を生成し、
発話提示部が、前記システム発話をユーザへ提示し、
発話受付部が、前記ユーザが発話したユーザ発話を受け付け、
対話制御部が、前記ユーザ発話に基づいて前記次に行う対話行為を決定する、
対話方法。 - 質問応答収集部が、発話とその発話を言い換えた発話とを含む言い替えデータを収集し、
変換モデル生成部が、前記言い替えデータを用いて、発話を入力とし、その発話を言い換えた発話を出力する発話変換モデルを学習し、
発話生成部が、システム発話を生成し、
発話変換部が、前記システム発話を前記発話変換モデルに入力して前記システム発話を言い換えた変換後システム発話を得、
発話提示部が、前記変換後システム発話をユーザへ提示する、
対話方法。 - 請求項1から4のいずれかに記載の対話装置としてコンピュータを機能させるためのプログラム。
Priority Applications (3)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| PCT/JP2021/019515 WO2022249221A1 (ja) | 2021-05-24 | 2021-05-24 | 対話装置、対話方法、およびプログラム |
| US18/562,294 US20240242718A1 (en) | 2021-05-24 | 2021-05-24 | Dialogue apparatus, dialogue method, and program |
| JP2023523706A JPWO2022249221A1 (ja) | 2021-05-24 | 2021-05-24 |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| PCT/JP2021/019515 WO2022249221A1 (ja) | 2021-05-24 | 2021-05-24 | 対話装置、対話方法、およびプログラム |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2022249221A1 true WO2022249221A1 (ja) | 2022-12-01 |
Family
ID=84229649
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/JP2021/019515 Ceased WO2022249221A1 (ja) | 2021-05-24 | 2021-05-24 | 対話装置、対話方法、およびプログラム |
Country Status (3)
| Country | Link |
|---|---|
| US (1) | US20240242718A1 (ja) |
| JP (1) | JPWO2022249221A1 (ja) |
| WO (1) | WO2022249221A1 (ja) |
Cited By (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2024110514A (ja) * | 2023-02-03 | 2024-08-16 | 日本特殊陶業株式会社 | バーチャルアシスタント装置、バーチャルアシスタントシステム、及びバーチャルアシスタント装置用のプログラム |
Families Citing this family (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN120929574B (zh) * | 2025-08-07 | 2026-03-06 | 武汉铁路职业技术学院 | 应用于大学生就业辅导的求职对话意图深度挖掘方法及系统 |
Citations (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2016126452A (ja) * | 2014-12-26 | 2016-07-11 | 株式会社小学館ミュージックアンドデジタルエンタテイメント | 会話処理ステム、会話処理方法、及び会話処理プログラム |
| JP2020190585A (ja) * | 2019-05-20 | 2020-11-26 | 日本電信電話株式会社 | 自動対話装置、自動対話方法、およびプログラム |
Family Cites Families (27)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US20030217135A1 (en) * | 2002-05-17 | 2003-11-20 | Masayuki Chatani | Dynamic player management |
| WO2006040971A1 (ja) * | 2004-10-08 | 2006-04-20 | Matsushita Electric Industrial Co., Ltd. | 対話支援装置 |
| JP2008203559A (ja) * | 2007-02-20 | 2008-09-04 | Toshiba Corp | 対話装置及び方法 |
| US10540976B2 (en) * | 2009-06-05 | 2020-01-21 | Apple Inc. | Contextual voice commands |
| WO2013155619A1 (en) * | 2012-04-20 | 2013-10-24 | Sam Pasupalak | Conversational agent |
| US9064491B2 (en) * | 2012-05-29 | 2015-06-23 | Nuance Communications, Inc. | Methods and apparatus for performing transformation techniques for data clustering and/or classification |
| CN103049433B (zh) * | 2012-12-11 | 2015-10-28 | 微梦创科网络科技(中国)有限公司 | 自动问答方法、自动问答系统及构建问答实例库的方法 |
| US9286892B2 (en) * | 2014-04-01 | 2016-03-15 | Google Inc. | Language modeling in speech recognition |
| JP6819988B2 (ja) * | 2016-07-28 | 2021-01-27 | 国立研究開発法人情報通信研究機構 | 音声対話装置、サーバ装置、音声対話方法、音声処理方法およびプログラム |
| US11183170B2 (en) * | 2016-08-17 | 2021-11-23 | Sony Corporation | Interaction control apparatus and method |
| JP6709997B2 (ja) * | 2016-09-23 | 2020-06-17 | パナソニックIpマネジメント株式会社 | 翻訳装置、翻訳システム、および評価サーバ |
| WO2019077897A1 (ja) * | 2017-10-17 | 2019-04-25 | ソニー株式会社 | 情報処理装置、情報処理方法、およびプログラム |
| US20190122661A1 (en) * | 2017-10-23 | 2019-04-25 | GM Global Technology Operations LLC | System and method to detect cues in conversational speech |
| US11688268B2 (en) * | 2018-01-23 | 2023-06-27 | Sony Corporation | Information processing apparatus and information processing method |
| US11381529B1 (en) * | 2018-12-20 | 2022-07-05 | Wells Fargo Bank, N.A. | Chat communication support assistants |
| CN110309850A (zh) * | 2019-05-15 | 2019-10-08 | 山东省计算中心(国家超级计算济南中心) | 基于语言先验问题识别和缓解的视觉问答预测方法及系统 |
| JP7491309B2 (ja) * | 2019-05-30 | 2024-05-28 | ソニーグループ株式会社 | 情報処理装置 |
| CN110609891B (zh) * | 2019-09-18 | 2021-06-08 | 合肥工业大学 | 一种基于上下文感知图神经网络的视觉对话生成方法 |
| JPWO2021106080A1 (ja) * | 2019-11-26 | 2021-06-03 | ||
| CN113392288A (zh) * | 2020-03-11 | 2021-09-14 | 阿里巴巴集团控股有限公司 | 视觉问答及其模型训练的方法、装置、设备及存储介质 |
| CN111460121B (zh) * | 2020-03-31 | 2022-07-08 | 思必驰科技股份有限公司 | 视觉语义对话方法及系统 |
| CN111460172A (zh) * | 2020-03-31 | 2020-07-28 | 北京小米移动软件有限公司 | 产品问题的答案确定方法、装置和电子设备 |
| US12002458B1 (en) * | 2020-09-04 | 2024-06-04 | Amazon Technologies, Inc. | Autonomously motile device with command processing |
| US20220093093A1 (en) * | 2020-09-21 | 2022-03-24 | Amazon Technologies, Inc. | Dialog management for multiple users |
| US11115353B1 (en) * | 2021-03-09 | 2021-09-07 | Drift.com, Inc. | Conversational bot interaction with utterance ranking |
| US20230177581A1 (en) * | 2021-12-03 | 2023-06-08 | Accenture Global Solutions Limited | Product metadata suggestion using embeddings |
| US20240412720A1 (en) * | 2023-06-11 | 2024-12-12 | Sergiy Vasylyev | Real-time contextually aware artificial intelligence (ai) assistant system and a method for providing a contextualized response to a user using ai |
-
2021
- 2021-05-24 US US18/562,294 patent/US20240242718A1/en active Pending
- 2021-05-24 JP JP2023523706A patent/JPWO2022249221A1/ja active Pending
- 2021-05-24 WO PCT/JP2021/019515 patent/WO2022249221A1/ja not_active Ceased
Patent Citations (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2016126452A (ja) * | 2014-12-26 | 2016-07-11 | 株式会社小学館ミュージックアンドデジタルエンタテイメント | 会話処理ステム、会話処理方法、及び会話処理プログラム |
| JP2020190585A (ja) * | 2019-05-20 | 2020-11-26 | 日本電信電話株式会社 | 自動対話装置、自動対話方法、およびプログラム |
Non-Patent Citations (2)
| Title |
|---|
| KOMACHI, MAMORU ET AL.: "Neural Paraphrase Generation for NLP for Education", JOURNAL OF THE JAPANESE SOCIETY FOR ARTIFICIAL INTELLIGENCE, vol. 34, no. 4, pages 451 - 459, ISSN: 2188-2266 * |
| TSUNOMORI, YUIKO ET AL.: "Development of a customizable open domain chat- oriented dialogue system", 84TH SPEECH AND LANGUAGE UNDERSTANDING AND DIALOGUE WORKSHOP, 15 November 2018 (2018-11-15), pages 124 - 127, ISSN: 0918-5682 * |
Cited By (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2024110514A (ja) * | 2023-02-03 | 2024-08-16 | 日本特殊陶業株式会社 | バーチャルアシスタント装置、バーチャルアシスタントシステム、及びバーチャルアシスタント装置用のプログラム |
| JP7675118B2 (ja) | 2023-02-03 | 2025-05-12 | 日本特殊陶業株式会社 | バーチャルアシスタント装置、バーチャルアシスタントシステム、及びバーチャルアシスタント装置用のプログラム |
Also Published As
| Publication number | Publication date |
|---|---|
| US20240242718A1 (en) | 2024-07-18 |
| JPWO2022249221A1 (ja) | 2022-12-01 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US11823061B2 (en) | Systems and methods for continual updating of response generation by an artificial intelligence chatbot | |
| CN101189659B (zh) | 用于认知超负荷的设备用户的交互式对话 | |
| CN113128239A (zh) | 促进以多种语言与自动化助理的端到端沟通 | |
| McTear et al. | Voice application development for Android | |
| Wu et al. | Research on business English translation framework based on speech recognition and wireless communication | |
| US20230026945A1 (en) | Virtual Conversational Agent | |
| Ouaddi et al. | DSL‐Driven Approaches and Metamodels for Chatbot Development: A Systematic Literature Review | |
| US20240242718A1 (en) | Dialogue apparatus, dialogue method, and program | |
| CN116795970A (zh) | 一种对话生成方法及其在情感陪护中的应用 | |
| JP7755504B2 (ja) | 開発支援システム及び開発支援方法 | |
| Bharathi et al. | Multi-Dialect Speech Corpus Creation for Enhancing Tamil Automatic Speech Recognition | |
| Patil | Healthcare chatbot using artificial intelligence | |
| JP2023159261A (ja) | 情報処理システム、情報処理方法及びプログラム | |
| CN115905475A (zh) | 答案评分方法、模型训练方法、装置、存储介质及设备 | |
| de Vries et al. | “You Can Do It!”—Crowdsourcing Motivational Speech and Text Messages | |
| Li et al. | Speech interaction of educational robot based on Ekho and Sphinx | |
| Lin | Reinforcement Learning in spoken language understanding (SLU): giving machines an ear for understanding | |
| Oliverio et al. | AIML+: Enhancing AIML for the Educational Domain Through Frames and Large Language Models | |
| US20240339041A1 (en) | Conversational teaching method and system and server thereof | |
| US20260010787A1 (en) | Artificial Intelligence Long-term Memory System | |
| WO2018147435A1 (ja) | 学習支援システム及び方法、並びにコンピュータプログラム | |
| Jerlina et al. | Introduction of the Rudiments of NLP: A Survey of Methods and Approaches to Natural Language Processing | |
| Rozga | Applying our learnings: Alexa skills kit | |
| Bhandari | Speech-To-Model: A Framework for Creating Software Models Using Voice Commands | |
| Papadopoulou | Developing and Evaluating a Chatbot |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 21942879 Country of ref document: EP Kind code of ref document: A1 |
|
| WWE | Wipo information: entry into national phase |
Ref document number: 2023523706 Country of ref document: JP |
|
| WWE | Wipo information: entry into national phase |
Ref document number: 18562294 Country of ref document: US |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 21942879 Country of ref document: EP Kind code of ref document: A1 |