WO2022222841A1 - 信息展示方法、装置、电子设备和计算机可读介质 - Google Patents
信息展示方法、装置、电子设备和计算机可读介质 Download PDFInfo
- Publication number
- WO2022222841A1 WO2022222841A1 PCT/CN2022/086827 CN2022086827W WO2022222841A1 WO 2022222841 A1 WO2022222841 A1 WO 2022222841A1 CN 2022086827 W CN2022086827 W CN 2022086827W WO 2022222841 A1 WO2022222841 A1 WO 2022222841A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- information
- displayed
- target
- user
- target user
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Images
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06Q—INFORMATION AND COMMUNICATION TECHNOLOGY [ICT] SPECIALLY ADAPTED FOR ADMINISTRATIVE, COMMERCIAL, FINANCIAL, MANAGERIAL OR SUPERVISORY PURPOSES; SYSTEMS OR METHODS SPECIALLY ADAPTED FOR ADMINISTRATIVE, COMMERCIAL, FINANCIAL, MANAGERIAL OR SUPERVISORY PURPOSES, NOT OTHERWISE PROVIDED FOR
- G06Q30/00—Commerce
- G06Q30/02—Marketing; Price estimation or determination; Fundraising
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L15/00—Speech recognition
- G10L15/08—Speech classification or search
- G10L15/18—Speech classification or search using natural language modelling
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L15/00—Speech recognition
- G10L15/22—Procedures used during a speech recognition process, e.g. man-machine dialogue
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L21/00—Speech or voice signal processing techniques to produce another audible or non-audible signal, e.g. visual or tactile, in order to modify its quality or its intelligibility
- G10L21/02—Speech enhancement, e.g. noise reduction or echo cancellation
- G10L21/0208—Noise filtering
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
- G10L25/48—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 specially adapted for particular use
- G10L25/51—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 specially adapted for particular use for comparison or discrimination
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10L—SPEECH ANALYSIS TECHNIQUES OR SPEECH SYNTHESIS; SPEECH RECOGNITION; SPEECH OR VOICE PROCESSING TECHNIQUES; SPEECH OR AUDIO CODING OR DECODING
- G10L25/00—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00
- G10L25/48—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 specially adapted for particular use
- G10L25/51—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 specially adapted for particular use for comparison or discrimination
- G10L25/63—Speech or voice analysis techniques not restricted to a single one of groups G10L15/00 - G10L21/00 specially adapted for particular use for comparison or discrimination for estimating an emotional state
Definitions
- Embodiments of the present disclosure relate to the field of computer technology, and in particular, to information display methods, apparatuses, electronic devices, and computer-readable media.
- voice interaction is the current mainstream interaction method.
- Many companies have also launched a variety of voice interaction products (for example, smart speakers, etc.).
- voice interaction products for example, smart speakers, etc.
- the existing voice interactive conversation method is often used: first, the information to be displayed is imported as content, and the user is directly asked relevant questions. Then, when the user gives voice feedback to the question, provide brand, product information related to the information to be displayed or launch an application program that meets the user's needs.
- Some embodiments of the present disclosure propose information display methods, apparatuses, electronic devices, and computer-readable media to solve one or more of the technical problems mentioned in the background section above.
- some embodiments of the present disclosure provide an information display method, including: acquiring audio data related to a target user; generating emotional information of the target user in response to determining that the audio data is not noise audio data; The type of information to be displayed corresponding to the above emotional information, wherein the type of information to be displayed represents the willingness of the user to receive the information to be displayed; in response to determining that the type of information to be displayed is not the target type of information to be displayed, determine the type to be displayed to the target user
- the first target information to be displayed is displayed to the above target user.
- some embodiments of the present disclosure provide an information display apparatus, comprising: an acquisition unit configured to acquire audio data related to a target user; a generation unit configured to respond to determining that the audio data is not noise audio data to generate the emotional information of the target user; the first determination unit is configured to determine the type of information to be displayed corresponding to the emotional information, wherein the type of the information to be displayed represents the willingness of the user to receive the information to be displayed; the second The determining unit is configured to, in response to determining that the type of information to be displayed is not the type of information to be displayed, to determine the first target information to be displayed to be displayed to the target user, the prompt corresponding to the type of information to be displayed, and the value of the prompt. Play the tone; the display unit is configured to play the prompt to the target user according to the play tone, and display the first target information to be displayed to the target user in response to the end of the prompt.
- some embodiments of the present disclosure provide an electronic device, including: at least one processor; and a storage device on which at least one program is stored, when the at least one program is executed by the at least one processor, the at least one process
- the device implements a method as described in any one of the implementations of the first aspect.
- some embodiments of the present disclosure provide a computer-readable medium having a computer program stored thereon, wherein the program, when executed by a processor, implements the method described in any implementation manner of the first aspect.
- FIG. 1 is a schematic diagram of an application scenario of an information display method according to some embodiments of the present disclosure
- FIG. 3 is a schematic diagram of obtaining a user emotional portrait in some embodiments of the information display method according to the present disclosure
- FIG. 4 is a flowchart of other embodiments of the information display method according to the present disclosure.
- FIG. 5 is a schematic structural diagram of some embodiments of an information display apparatus according to the present disclosure.
- FIG. 6 is a schematic structural diagram of an electronic device suitable for implementing some embodiments of the present disclosure.
- some embodiments of the present disclosure provide information display methods and apparatuses, which can quickly, efficiently, and more specifically push the information to be displayed to users for browsing, thereby improving user experience.
- FIG. 1 is a schematic diagram of an application scenario of an information display method according to some embodiments of the present disclosure.
- the electronic device 101 may first acquire audio data 103 related to the target user 102 . Then, in response to determining that the above-mentioned audio data 103 is not noise audio data, the above-mentioned emotion information 104 of the target user 102 is generated. In this application scenario, the above-mentioned emotional information 104 may be: "happy". Further, the type of information to be displayed 105 corresponding to the above-mentioned emotional information 104 is determined. The above information type 105 to be displayed represents the willingness of the user to receive the information to be displayed. In this application scenario, the above-mentioned type of information to be displayed 105 may be: "positive type of information to be displayed".
- the first target to-be-displayed information 106 to be displayed to the above-mentioned target user 102 , the prompt 107 corresponding to the above-mentioned type of information to be displayed 105 , and the above-mentioned prompt 107 of the playing intonation 108.
- the above-mentioned first target information to be displayed 106 may be: "beautiful style, pure care”.
- the above prompt 107 may be: "I am in a good mood today, I recommend a book to you!.
- the above-mentioned playing tone 108 may be: "positive, sunny”.
- the prompt 107 is played to the target user 102, and in response to the end of the prompt 107, the first target to-be-displayed information 106 is displayed to the target user 102.
- the above electronic device 101 may be hardware or software.
- the electronic device When the electronic device is hardware, it can be implemented as a distributed cluster composed of multiple servers or terminal devices, or can be implemented as a single server or a single terminal device.
- the electronic device When the electronic device is embodied as software, it can be installed in the hardware devices listed above. It can be implemented, for example, as multiple software or software modules for providing distributed services, or as a single software or software module. There is no specific limitation here.
- FIG. 1 the number of electronic devices in FIG. 1 is merely illustrative. There may be any number of electronic devices depending on implementation needs.
- the information display method includes the following steps:
- Step 201 Acquire audio data related to the target user.
- the execution body of the above information display method may acquire and collect audio data related to the target user by using a related recording device.
- the above-mentioned audio data may be sound data of the environment where the target user is recorded.
- the above-mentioned execution subject may have a conversation with the above-mentioned target user to acquire audio data related to the above-mentioned target user.
- the above-mentioned conversation with the above-mentioned target user to obtain audio data related to the above-mentioned target user may include the following steps:
- the above-mentioned executive body may instruct the above-mentioned target user to complete the required user behavior and/or reply to the proposed target question.
- the above-mentioned required user behavior may be to instruct the target user to check in the target application by means of a voice reply or manual operation.
- the target question raised above may be: "How are you feeling today?".
- the above target application may be a mobile terminal application (APP, Application).
- the executive body may receive audio data generated by the target user performing the required user behavior and/or responding to the proposed target question.
- the above-mentioned conversation with the above-mentioned target user to obtain audio data related to the above-mentioned target user may include the following steps:
- the above target user is instructed to give voice feedback to the information to be displayed on the second target.
- the audio data related to the information to be displayed on the second target is preset and randomly inserted into the target audio.
- the above-mentioned target audio may be the audio presented by the target application.
- the second step is to receive the audio data related to the above target user.
- the audio data related to the information to be displayed on the second target is determined through the following steps.
- the above target application may be a mobile terminal application (APP, Application).
- the execution body may determine whether the target user is a user who uses the target application for the first time by querying a user information database.
- the above user information database stores the login information of the target user.
- the second step in response to determining that it is not, obtain a pre-stored user emotional portrait associated with the target user.
- the above-mentioned user emotional portrait stores historical emotional information of the target user.
- the above-mentioned execution subject may acquire a pre-stored user emotional portrait associated with the above-mentioned target user from a user emotional portrait database.
- the third step is to determine the second target information to be displayed in each pre-stored information to be displayed according to the above-mentioned user emotional portrait.
- the above-mentioned execution subject may determine, according to the above-mentioned user emotion portrait, a certain type of emotion corresponding to the target user that occurs most frequently. Then, the second target information to be displayed that most frequently matches a certain type of emotion is determined from the above-mentioned pieces of information to be displayed.
- the above-mentioned execution subject may acquire the above-mentioned user emotional portrait in response to receiving the target authorization signal.
- the target authorization signal may be a signal generated by a user corresponding to the user portrait performing a target operation on the target control.
- the above target controls can be included in the authorization prompt box.
- the above authorization prompt box can be displayed on the target terminal device.
- the target terminal device may be a terminal device logged in with the account corresponding to the user.
- the above-mentioned terminal device may be a "mobile phone” or a "computer”.
- the above target operation may be a "click operation” or a "slide operation”.
- the aforementioned target control may be a "confirm button".
- the above authorization prompt box may be as shown in FIG. 3 .
- the above authorization prompt box may include: a prompt information display part 301 and a control 302 .
- the above-mentioned prompt information display part 301 can be used for displaying prompt information.
- the above prompt information may be "whether it is allowed to obtain the user's emotional portrait”.
- the above-mentioned control 302 may be a "confirm button” or a "cancel button”.
- the fourth step is to match the audio data to be played that is related to the information to be displayed on the second target.
- the above-mentioned execution body may generate audio data to be played that is related to the above-mentioned second target information to be displayed by matching a preset template.
- the above-mentioned second information to be displayed may be: "infinite detachment, bit by bit creation”.
- the corresponding audio data can be: Do you need to know the content information related to "infinite detachment, bit by bit creation”? .
- Step 202 in response to determining that the audio data is not noise audio data, generate emotional information of the target user.
- the execution subject in response to determining that the audio data is not noise audio data, may generate the emotion information of the target user.
- the above-mentioned emotional information may include, but is not limited to, at least one of the following: happy, surprised, sad, afraid, angry, and disgusted.
- the foregoing execution subject may generate the emotion information of the foregoing target user through the following steps:
- noise reduction is performed on the above audio data to obtain the audio data after noise reduction.
- the above-mentioned execution body may use a noise removal algorithm to perform noise reduction on the above-mentioned audio data, and may obtain the de-noised audio data.
- the above-mentioned noise removal algorithm may be a minimum controlled recursive averaging algorithm (MCRA).
- the second step is to use the natural language processing method to fuse the statistical characteristics and time series characteristics of acoustic parameters to classify the denoised audio data to obtain the emotional information of the target user.
- the acoustic parameters can be obtained through the following steps:
- the first step is to obtain the prosody parameters of the audio.
- the prosody parameters include: fundamental frequency parameters and duration parameters.
- the fundamental frequency parameter can be obtained by the YIN algorithm, and the duration parameter can be obtained by using a related labeling tool.
- the second step is to obtain the spectral parameters of the audio.
- the spectrum parameters include: Mel spectrum cepstrum parameter MFCC (Mel Frequency Cepstrum Coefficient), spectrum centroid parameter Sc, spectrum cutoff parameter Sr, spectrum transition parameter Sf, and frequency band periodic parameter Sp.
- A(n) can be the amplitude corresponding to the nth spectral line
- the calculation formula of Sc can be:
- the formula for calculating Sr can be:
- Ai(n) and Ai-1(n) as the amplitude spectrum of the current frame and the previous frame, respectively, and the S formula for calculating Sf can be:
- the third step is to process the statistical features of the above parameters in the audio data, and use a probabilistic neural network (PNN, Product Network) to identify the statistical features.
- PNN probabilistic neural network
- the fourth step is to process the time series features of the above parameters in the audio data, and use a Hidden Markov Model (HMM, Hidden Markov Model) to identify the time series features.
- HMM Hidden Markov Model
- the fifth step is to extract N groups of features (including statistical and time series features) from the sample x, denoted as f1 ⁇ fN.
- the probability of belonging to the i-th emotion obtained by the PNN or HMM model is: P(C i
- F represents the fusion rule, and the recognition results of the two features are fused according to the "multiplication principle” and the “addition principle". According to existing research, this algorithm performs better in reducing data confusion.
- the multiplication rule formula can be:
- the above steps further include:
- the emotional information of the target user is integrated into the user emotional portrait associated with the target user.
- the accumulation of the user's emotional portraits associated with the above-mentioned target users can make the emotional changes in the user's life more clear.
- information to be displayed can be provided for the target user according to the user emotional portrait of the target user in a more targeted manner.
- Step 203 Determine the type of information to be displayed corresponding to the above emotional information.
- the above-mentioned execution body may determine the type of information to be displayed corresponding to the above-mentioned emotional information.
- the types of the information to be displayed may include: a positive type of the information to be displayed, an empathy type of the information to be displayed, and a rejection type of the information to be displayed.
- the above type of information to be displayed represents the willingness of the user to receive the information to be displayed.
- the above-mentioned active type of the information to be displayed can represent that the user is more willing to receive the information to be displayed.
- the above-mentioned rejection type of the information to be displayed may indicate that the user is less willing to receive the information to be displayed.
- the type of information to be displayed for which it is unclear whether the user is willing to receive the information to be displayed may be determined as the empathy type of information to be displayed.
- the emotional information and the type of information to be displayed may be in a one-to-one correspondence.
- the corresponding emotional information may include: happy, surprised, and neutral.
- the corresponding emotional information may include: anger, disgust.
- the corresponding emotional information may include: sadness and fear.
- the above-mentioned execution subject may determine the type of information to be displayed corresponding to the above-mentioned emotional information by using a pre-built relationship table between emotional information and the type of information to be displayed.
- Step 204 in response to determining that the above-mentioned type of information to be displayed is not the type of target information to be displayed, determine the above-mentioned first target to-be-displayed information to be displayed to the above-mentioned target user, a prompt corresponding to the above-mentioned type of information to be displayed, and the playing tone of the above-mentioned prompt.
- the execution body may determine the first target information to be displayed and a prompt corresponding to the type of the information to be displayed to be displayed to the target user and the playing intonation of the above prompt.
- the above target information type to be displayed may be a rejection type of information to be displayed. Because when the type of information to be displayed is the rejection type of information to be displayed, the emotional information representing the target user may be anger and disgust. Therefore, the above-mentioned execution body can selectively push the information to be displayed on the first target.
- the first target information to be displayed corresponding to the type of information to be displayed exists may represent preset information to be displayed that has a corresponding relationship with the type of information to be displayed.
- the positive type of the information to be displayed corresponds to the information to be displayed that can keep the user in a pleasant mood.
- the information to be displayed may be information to be displayed in the category of cosmetics and skin care.
- the rejection type of the information to be displayed corresponds to the information to be displayed that allows the user to eliminate negative emotions.
- the information to be displayed may be information to be displayed of storage articles.
- the empathy type of the information to be displayed corresponds to the information to be displayed that allows the user to relieve tension.
- the information to be displayed may be information to be displayed of snack foods.
- the prompt corresponding to the type of information to be displayed may be preset.
- the prompt corresponding to the positive type of information to be displayed can be: "I am in a good mood today, I recommend a book to you!.
- the prompt corresponding to the empathy type of the information to be displayed can be: "I am in a normal mood today, let's take a look at the surprising little accessories ⁇ ”.
- the prompt corresponding to the rejection type of the information to be displayed may be: "Thank you”.
- the playing tone of the prompt corresponding to the type of information to be displayed may be preset.
- the playing tone corresponding to the positive type of the information to be displayed may be a positive and sunny playing tone.
- the prompt corresponding to the empathy type of the information to be displayed may be a peaceful and gentle tone of play.
- the prompt corresponding to the rejection type of the information to be displayed may be a playing tone with a more formal tone.
- Step 205 according to the above-mentioned playing tone, playing the above-mentioned prompt language to the above-mentioned target user, and in response to the end of playing the above-mentioned prompt language, displaying the above-mentioned first target to-be-displayed information to the above-mentioned target user.
- the execution body may play the prompt to the target user according to the playing tone, and display the information to be displayed on the first target to the target user in response to the end of playing the prompt.
- the information display methods of some embodiments of the present disclosure can quickly, efficiently and more specifically push the information to be displayed to the user for the user to browse, thereby improving the user experience. Specifically, it cannot efficiently and accurately determine whether the current user wishes to view the information to be displayed. Specifically, in a situation where it is uncertain whether the user wishes to view the information to be displayed, displaying the information to be displayed to the target user may result in poor user experience. Based on this, the information display methods of some embodiments of the present disclosure may first acquire audio data related to the target user for subsequent determination of emotional information of the current target user. Then, in response to determining that the audio data is not noise audio data, the emotion information of the target user is generated.
- the willingness of the current user to receive the information to be displayed is further determined by generating emotional information of the target user. Further, the type of information to be displayed corresponding to the above emotional information is determined.
- the above-mentioned type of information to be displayed represents the willingness of the user to receive the information to be displayed.
- the user's emotional information may include various types. It may be cumbersome and complicated to directly determine whether to push the information to be displayed and/or what kind of information to be displayed to be pushed subsequently through the user's emotional information. Thus, by determining the type of information to be displayed corresponding to the emotional information, the willingness of the target user can be more specifically and clearly determined.
- the first target to-be-displayed information to be displayed to the above-mentioned target user that the current target user may wish to see and a prompt corresponding to the above-mentioned type of information to be displayed and the playing intonation of the above prompt.
- the type of the target information to be displayed may indicate that the target user has no great willingness to receive the information to be displayed.
- the information display method fully considers the emotional information of the target user when displaying the information to be displayed, so that the target user can view the to-be-displayed information that may meet the needs of the current user.
- the side In order to achieve the effect of displaying the information to be displayed, the side also greatly improves the user experience.
- the information display method includes the following steps:
- Step 401 Acquire audio data related to the target user.
- Step 402 in response to determining that the audio data is not noise audio data, generate emotional information of the target user.
- Step 403 Determine the type of information to be displayed corresponding to the above emotional information.
- Step 404 In response to determining that the type of information to be displayed is not the type of information to be displayed, determine the first target information to be displayed to be displayed to the target user, a prompt corresponding to the type of information to be displayed, and the playing tone of the prompt.
- Step 405 playing the prompt language to the target user according to the playing tone, and displaying the information to be displayed on the first target to the target user in response to the end of playing the prompt language.
- steps 401-405 for the specific implementation of steps 401-405 and the technical effects brought about by them, reference may be made to steps 201-205 in the embodiment corresponding to FIG. 2, and details are not repeated here.
- Step 406 in response to determining that the audio data is the noise audio data, obtain user behavior information of the target user.
- the execution subject in response to determining that the audio data is the noise audio data, may obtain the user behavior information of the target user by querying a user information database.
- the above-mentioned user behavior information may include, but is not limited to, at least one of the following: the user's page click information, the user's page browsing information, the information that the user performs the first value transfer operation (purchase), the user's second value transfer operation (transfer) )Information.
- the above-mentioned execution body may determine whether the audio data is the above-mentioned noise audio data through a related natural language processing algorithm.
- Step 407 Determine the type of information to be displayed corresponding to the above-mentioned user behavior information.
- the above-mentioned execution body may determine the type of information to be displayed corresponding to the above-mentioned user behavior information.
- the above-mentioned executive body may determine the type of information to be displayed corresponding to the above-mentioned user behavior information by using a pre-established table of association relationships between the user behavior information and the type of information to be displayed.
- the above-mentioned execution body can count the occurrences of various user behavior information. Then, the number of occurrences of each user behavior information corresponding to each type of information to be displayed is determined by using the above table. Finally, the type of information to be displayed with the most occurrences of each type of information to be displayed is determined as the type of information to be displayed corresponding to the above-mentioned user behavior information.
- Step 408 in response to determining that the above-mentioned type of information to be displayed is not the above-mentioned target type of information to be displayed, determine the third target to-be-displayed information associated with the above-mentioned type of information to be displayed to be displayed to the above-mentioned target user, and the above-mentioned type of information to be displayed corresponds to the type of information to be displayed.
- the prompt words and the playing intonation of the above prompt words in response to determining that the above-mentioned type of information to be displayed is not the above-mentioned target type of information to be displayed.
- Step 409 according to the playing tone, play the prompt language to the target user, and in response to the end of the prompt language playing, display the third target information to be displayed to the target user.
- steps 408-409 for the specific implementation of steps 408-409 and the technical effects brought about by them, reference may be made to steps 204-205 in the embodiment corresponding to FIG. 2, and details are not repeated here.
- the process 400 of the information display method in some embodiments corresponding to FIG. 4 more highlights the specific details of displaying the information to be displayed according to the user behavior information step. Therefore, the solutions described in these embodiments can analyze the user behavior information, and push the information to be displayed more targetedly and efficiently without the need to obtain effective user audio data in real time, thereby improving the user experience.
- the present disclosure provides some embodiments of an information display apparatus. These apparatus embodiments correspond to those method embodiments shown in FIG. 2 , and the apparatus may specifically be Used in various electronic devices.
- an information display apparatus 500 includes: an acquisition unit 501 , a generation unit 502 , a first determination unit 503 , a second determination unit 504 and a display unit 505 .
- the acquiring unit 501 is configured to acquire audio data related to the target user;
- the generating unit 502 is configured to generate emotional information of the target user in response to determining that the audio data is not noise audio data;
- the first determining unit 503, is configured to determine the type of information to be displayed corresponding to the above-mentioned emotional information, wherein the above-mentioned type of information to be displayed represents the willingness of the user to receive the information to be displayed;
- the second determining unit 504 is configured to respond to determining the above-mentioned type of information to be displayed Not the type of the target information to be displayed, determine the first target information to be displayed to be displayed to the target user, the prompt corresponding to the above information type to be displayed, and the playing tone of the prompt;
- the display unit 505 is configured to be based on the
- playing the above-mentioned prompt words to the above-mentioned target users and in response to the end of playing the above-mentioned prompt words, displaying the above-mentioned first target information to be displayed to the above-mentioned target users.
- the above-mentioned obtaining unit 501 may be configured to: have a conversation with the above-mentioned target user to obtain audio data related to the above-mentioned target user.
- the above-mentioned obtaining unit 501 may be configured to: instruct the above-mentioned target user to complete the required user behavior and/or reply to the proposed target question; receive the above-mentioned target user to perform the required Audio data resulting from user actions and/or responses to targeted questions posed.
- the obtaining unit 501 may be configured to: instruct the target user to give voice feedback to the information to be displayed on the second target, wherein the audio related to the information to be displayed on the second target
- the data is preset audio data randomly inserted into the target audio; the audio data related to the target user is received.
- the above-mentioned apparatus 500 further includes: a user information acquisition unit, a third determination unit, a fourth determination unit, and a playback display unit (not shown in the figure).
- the user information obtaining unit may be configured to: in response to determining that the audio data is the noise audio data, obtain the user behavior information of the target user.
- the third determining unit may be configured to: determine the type of information to be displayed corresponding to the above-mentioned user behavior information.
- the fourth determination unit may be configured to: in response to determining that the above-mentioned type of information to be displayed is not the above-mentioned target type of information to be displayed, determine the third target information to be displayed, which is associated with the above-mentioned type of information to be displayed, to be displayed to the above-mentioned target user, The prompt corresponding to the type of information to be displayed and the playing tone of the prompt.
- the playing and presenting unit may be configured to: play the prompt word to the target user according to the play tone, and display the third target information to be displayed to the target user in response to the end of the play of the prompt word.
- the above-mentioned apparatus 500 further includes: a playing unit (not shown in the figure).
- the playing unit may be configured to: in response to determining that the above-mentioned type of information to be displayed is the above-mentioned target type of information to be displayed, use the target intonation to play the conclusion.
- the audio data related to the information to be displayed on the second target is determined by the following steps: in response to determining that the target user has logged in to the target application, determining whether the target user is the first target user A user who uses the above-mentioned target application once; in response to determining that it is not, obtain a pre-stored user emotional portrait associated with the above-mentioned target user; according to the above-mentioned user emotional portrait, determine the second target to be displayed in each pre-stored information to be displayed information; matches the audio data to be played that is related to the information to be displayed on the second target.
- the above-mentioned apparatus 500 further includes: an integration unit (not shown in the figure).
- the integration unit may be configured to: integrate the emotional information of the target user into the user emotional portrait associated with the target user.
- the units recorded in the apparatus 500 correspond to the respective steps in the method described with reference to FIG. 2 . Therefore, the operations, features and beneficial effects described above with respect to the method are also applicable to the apparatus 500 and the units included therein, and details are not described herein again.
- FIG. 6 there is shown a schematic structural diagram of an electronic device (eg, the electronic device in FIG. 1 ) 600 suitable for implementing some embodiments of the present disclosure.
- the electronic device shown in FIG. 6 is only an example, and should not impose any limitation on the function and scope of use of the embodiments of the present disclosure.
- an electronic device 600 may include a processing device (eg, a central processing unit, a graphics processor, etc.) 601 that may be loaded into random access according to a program stored in a read only memory (ROM) 602 or from a storage device 608 Various appropriate actions and processes are executed by the programs in the memory (RAM) 603 . In the RAM 603, various programs and data required for the operation of the electronic device 600 are also stored.
- the processing device 601, the ROM 602, and the RAM 603 are connected to each other through a bus 604.
- An input/output (I/O) interface 605 is also connected to bus 604 .
- I/O interface 605 input devices 606 including, for example, a touch screen, touchpad, keyboard, mouse, camera, microphone, accelerometer, gyroscope, etc.; including, for example, a liquid crystal display (LCD), speakers, vibration An output device 607 of a computer, etc.; a storage device 608 including, for example, a magnetic tape, a hard disk, etc.; and a communication device 609.
- Communication means 609 may allow electronic device 600 to communicate wirelessly or by wire with other devices to exchange data. While FIG. 6 shows electronic device 600 having various means, it should be understood that not all of the illustrated means are required to be implemented or provided. More or fewer devices may alternatively be implemented or provided. Each block shown in FIG. 6 may represent one device, or may represent multiple devices as required.
- the processes described above with reference to the flowcharts may be implemented as computer software programs.
- some embodiments of the present disclosure include a computer program product comprising a computer program carried on a computer-readable medium, the computer program containing program code for performing the method illustrated in the flowchart.
- the computer program may be downloaded and installed from the network via the communication device 609, or from the storage device 608, or from the ROM 602.
- the processing device 601 When the computer program is executed by the processing device 601, the above-mentioned functions defined in the methods of some embodiments of the present disclosure are performed.
- the computer-readable medium described above may be a computer-readable signal medium or a computer-readable storage medium, or any combination of the foregoing two.
- the computer-readable storage medium can be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus or device, or a combination of any of the above.
- a computer-readable storage medium can be any tangible medium that contains or stores a program that can be used by or in conjunction with an instruction execution system, apparatus, or device.
- a computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave, carrying computer-readable program code therein.
- Such propagated data signals may take a variety of forms including, but not limited to, electromagnetic signals, optical signals, or any suitable combination of the foregoing.
- a computer-readable signal medium can also be any computer-readable medium other than a computer-readable storage medium that can transmit, propagate, or transport the program for use by or in connection with the instruction execution system, apparatus, or device .
- Program code embodied on a computer readable medium may be transmitted using any suitable medium including, but not limited to, electrical wire, optical fiber cable, RF (radio frequency), etc., or any suitable combination of the foregoing.
- the client and server can use any currently known or future developed network protocol such as HTTP (HyperText Transfer Protocol) to communicate, and can communicate with digital data in any form or medium Communication (eg, a communication network) interconnects.
- HTTP HyperText Transfer Protocol
- Examples of communication networks include local area networks (“LAN”), wide area networks (“WAN”), the Internet (eg, the Internet), and peer-to-peer networks (eg, ad hoc peer-to-peer networks), as well as any currently known or future development network of.
- the above-mentioned computer-readable medium may be included in the above-mentioned electronic device; or may exist alone without being assembled into the electronic device.
- the above-mentioned computer-readable medium carries at least one program, and when the above-mentioned at least one program is executed by the electronic device, makes the electronic device: acquire audio data related to the target user; in response to determining that the above-mentioned audio data is not noise audio data, generate the above-mentioned audio data.
- the type of information to be displayed corresponding to the above emotional information, wherein the above type of information to be displayed represents the willingness of the user to receive the information to be displayed; in response to determining that the type of information to be displayed is not the target type of information to be displayed , determine the first target to-be-displayed information to be displayed to the above-mentioned target user, the prompt language corresponding to the type of the above-mentioned information to be displayed, and the playing tone of the above-mentioned prompt; When the above-mentioned prompt is played, the above-mentioned first target information to be displayed is displayed to the above-mentioned target user.
- Computer program code for carrying out operations of some embodiments of the present disclosure may be written in one or more programming languages, including object-oriented programming languages—such as Java, Smalltalk, C++, or a combination thereof, Also included are conventional procedural programming languages - such as the "C" language or similar programming languages.
- the program code may execute entirely on the user's computer, partly on the user's computer, as a stand-alone software package, partly on the user's computer and partly on a remote computer, or entirely on the remote computer or server.
- the remote computer may be connected to the user's computer through any kind of network, including a local area network (LAN) or a wide area network (WAN), or may be connected to an external computer (eg, using an Internet service provider to via Internet connection).
- LAN local area network
- WAN wide area network
- each block in the flowchart or block diagrams may represent a module, segment, or portion of code that contains at least one configurable function for implementing the specified logical function. Execute the instruction.
- the functions noted in the blocks may occur out of the order noted in the figures. For example, two blocks shown in succession may, in fact, be executed substantially concurrently, or the blocks may sometimes be executed in the reverse order, depending upon the functionality involved.
- each block of the block diagrams and/or flowchart illustrations, and combinations of blocks in the block diagrams and/or flowchart illustrations can be implemented in dedicated hardware-based systems that perform the specified functions or operations , or can be implemented in a combination of dedicated hardware and computer instructions.
- the units described in some embodiments of the present disclosure may be implemented by means of software, and may also be implemented by means of hardware.
- the described unit may also be provided in the processor, for example, it may be described as: a processor includes an acquisition unit, a generation unit, a first determination unit, a second determination unit and a display unit. Wherein, the names of these units do not constitute a limitation of the unit itself in some cases, for example, the acquisition unit may also be described as "a unit for acquiring audio data related to the target user".
- exemplary types of hardware logic components include: Field Programmable Gate Arrays (FPGAs), Application Specific Integrated Circuits (ASICs), Application Specific Standard Products (ASSPs), Systems on Chips (SOCs), Complex Programmable Logical Devices (CPLDs) and more.
- FPGAs Field Programmable Gate Arrays
- ASICs Application Specific Integrated Circuits
- ASSPs Application Specific Standard Products
- SOCs Systems on Chips
- CPLDs Complex Programmable Logical Devices
Landscapes
- Engineering & Computer Science (AREA)
- Health & Medical Sciences (AREA)
- Physics & Mathematics (AREA)
- Multimedia (AREA)
- Computational Linguistics (AREA)
- Acoustics & Sound (AREA)
- Human Computer Interaction (AREA)
- Audiology, Speech & Language Pathology (AREA)
- Signal Processing (AREA)
- Business, Economics & Management (AREA)
- Development Economics (AREA)
- Strategic Management (AREA)
- Artificial Intelligence (AREA)
- Psychiatry (AREA)
- Hospice & Palliative Care (AREA)
- Quality & Reliability (AREA)
- General Health & Medical Sciences (AREA)
- Accounting & Taxation (AREA)
- Child & Adolescent Psychology (AREA)
- Finance (AREA)
- Game Theory and Decision Science (AREA)
- General Business, Economics & Management (AREA)
- Economics (AREA)
- Marketing (AREA)
- Entrepreneurship & Innovation (AREA)
- Theoretical Computer Science (AREA)
- General Physics & Mathematics (AREA)
- User Interface Of Digital Computer (AREA)
Abstract
本公开的实施例公开了信息展示方法、装置、电子设备和介质。该方法的一具体实施方式包括:获取与目标用户相关的音频数据;响应于确定该音频数据不是噪声音频数据,生成该目标用户的情绪信息;确定与情绪信息相对应的待展示信息类型,其中,待展示信息类型表征用户接收待展示信息的意愿程度;响应于确定待展示信息类型不是目标待展示信息类型,确定待展示给目标用户的第一目标待展示信息、待展示信息类型对应的提示语和该提示语的播放语调;依照播放语调,播放提示语给目标用户,以及响应于提示语播放结束,将第一目标待展示信息展示给目标用户。该实施方式可以快捷、高效、更有针对性的将待展示信息推送给用户以供用户浏览,提高了用户体验。
Description
相关申请的交叉引用
本申请要求于2021年04月20日提交的,申请号为202110426791.4、发明名称为“信息展示方法、装置、电子设备和计算机可读介质”的中国专利申请的优先权,其全部内容作为整体并入本申请中。
本公开的实施例涉及计算机技术领域,具体涉及信息展示方法、装置、电子设备和计算机可读介质。
目前,语音交互是当前主流的交互方式。许多公司也相应推出了各式各样的语音交互产品(例如,智能音箱等)。对于待展示信息展现给用户,现有常常采用的语音交互会话方式:首先,将待展示信息作为内容导入,直接向用户提出相关问题。然后,当用户对该问题予以语音反馈时,提供与待展示信息相关品牌、产品信息或启动满足用户需要的应用程序。
发明内容
本公开的内容部分用于以简要的形式介绍构思,这些构思将在后面的具体实施方式部分被详细描述。本公开的内容部分并不旨在标识要求保护的技术方案的关键特征或必要特征,也不旨在用于限制所要求的保护的技术方案的范围。
本公开的一些实施例提出了信息展示方法、装置、电子设备和计算机可读介质,来解决以上背景技术部分提到的技术问题中的一项或多项。
第一方面,本公开的一些实施例提供了一种信息展示方法,包括:获取与目标用户相关的音频数据;响应于确定上述音频数据不是噪声音频数据,生成上述目标用户的情绪信息;确定与上述情绪信息相对应的待展示信息类型,其中,上述待展示信息类型表征用户接收待展示信息的意愿程度;响应于确定上述待展示信息类型不是目标待展示信息类型,确定待展示给上述目标用户的第一目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调;依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第一目标待展示信息展示给上述目标用户。
第二方面,本公开的一些实施例提供了一种信息展示装置,包括:获取单元,被配置成获取与目标用户相关的音频数据;生成单元,被配置成响应于确定上述音频数据不是噪声音频数据,生成上述目标用户的情绪信息;第一确定单元,被配置成确定与上述情绪信息相对应的待展示信息类型,其中,上述待展示信息类型表征用户接收待展示信息的意愿程度;第二确定单元,被配置成响应于确定上述待展示信息类型不是目标待展示信息类型,确定待展示给上述目标用户的第一目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调;展示单元,被配置成依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第一目标待展示信息展示给上述目标用户。
第三方面,本公开的一些实施例提供了一种电子设备,包括:至少一个处理器;存储装置,其上存储有至少一个程序,当至少一个程序被至少一个处理器执行,使得至少一个处理器实现如第一方面中任一实现方式描述的方法。
第四方面,本公开的一些实施例提供了一种计算机可读介质,其上存储有计算机程序,其中,程序被处理器执行时实现如第一方面中任一实现方式描述的方法。
结合附图并参考以下具体实施方式,本公开各实施例的上述和其 他特征、优点及方面将变得更加明显。贯穿附图中,相同或相似的附图标记表示相同或相似的元素。应当理解附图是示意性的,原件和元素不一定按照比例绘制。
图1是根据本公开的一些实施例的信息展示方法的一个应用场景的示意图;
图2是根据本公开的信息展示方法的一些实施例的流程图;
图3是根据本公开的信息展示方法的一些实施例中的获取用户情绪画像的示意图;
图4是根据本公开的信息展示方法的另一些实施例的流程图;
图5是根据本公开的信息展示装置的一些实施例的结构示意图;
图6是适于用来实现本公开的一些实施例的电子设备的结构示意图。
下面将参照附图更详细地描述本公开的实施例。虽然附图中显示了本公开的某些实施例,然而应当理解的是,本公开可以通过各种形式来实现,而且不应该被解释为限于这里阐述的实施例。相反,提供这些实施例是为了更加透彻和完整地理解本公开。应当理解的是,本公开的附图及实施例仅用于示例性作用,并非用于限制本公开的保护范围。
另外还需要说明的是,为了便于描述,附图中仅示出了与有关发明相关的部分。在不冲突的情况下,本公开中的实施例及实施例中的特征可以相互组合。
需要注意,本公开中提及的“第一”、“第二”等概念仅用于对不同的装置、模块或单元进行区分,并非用于限定这些装置、模块或单元所执行的功能的顺序或者相互依存关系。
需要注意,本公开中提及的“一个”、“多个”的修饰是示意性而非限制性的,本领域技术人员应当理解,除非在上下文另有明确指出,否则应该理解为“一个或多个”。
本公开实施方式中的多个装置之间所交互的消息或者信息的名称 仅用于说明性的目的,而并不是用于对这些消息或信息的范围进行限制。
相关的信息展示方法,例如,语音交互会话方式等经常会存在如下技术问题:不能高效、准确的确定出当前用户是否希望观看待展示信息。具体而言,在不确定用户是否希望观看待展示信息的情况下,展示待展示信息给目标用户,可能导致用户体验较差。
为了解决以上所阐述的问题,本公开的一些实施例提出了信息展示方法及装置,可以快捷、高效、更有针对性的将待展示信息推送给用户以供用户浏览,提高了用户体验。
下面将参考附图并结合实施例来详细说明本公开。
图1是根据本公开一些实施例的信息展示方法的一个应用场景的示意图。
在图1的应用场景中,电子设备101可以首先获取与目标用户102相关的音频数据103。然后,响应于确定上述音频数据103不是噪声音频数据,生成上述目标用户102的情绪信息104。在本应用场景中,上述情绪信息104可以是:“开心”。进而,确定与上述情绪信息104相对应的待展示信息类型105。其中,上述待展示信息类型105表征用户接收待展示信息的意愿程度。在本应用场景中,上述待展示信息类型105可以是:“待展示信息积极类型”。接着,响应于确定上述待展示信息类型105不是目标待展示信息类型,确定待展示给上述目标用户102的第一目标待展示信息106、上述待展示信息类型105对应的提示语107和上述提示语107的播放语调108。在本应用场景中,上述第一目标待展示信息106可以是:“亮丽风采,纯真关怀”。上述提示语107可以是:“今天心情不错,推荐一本书给你!”。上述播放语调108可以是:“积极、阳光”。最后,依照上述播放语调108,播放上述提示语107给上述目标用户102,以及响应于上述提示语107播放结束,将上述第一目标待展示信息106展示给上述目标用户102。
需要说明的是,上述电子设备101可以是硬件,也可以是软件。当电子设备为硬件时,可以实现成多个服务器或终端设备组成的分布 式集群,也可以实现成单个服务器或单个终端设备。当电子设备体现为软件时,可以安装在上述所列举的硬件设备中。其可以实现成例如用来提供分布式服务的多个软件或软件模块,也可以实现成单个软件或软件模块。在此不做具体限定。
应该理解,图1中的电子设备的数目仅仅是示意性的。根据实现需要,可以具有任意数目的电子设备。
继续参考图2,示出了根据本公开的信息展示方法的一些实施例的流程200。该信息展示方法,包括以下步骤:
步骤201,获取与目标用户相关的音频数据。
在一些实施例中,上述信息展示方法的执行主体(例如图1所示的电子设备)可以通过利用相关录音设备来获取收集与目标用户相关的音频数据。上述音频数据可以是录取目标用户所处环境的声音数据。
在一些实施例的一些可选的实现方式中,上述执行主体可以与上述目标用户会话以获取与上述目标用户相关的音频数据。
可选地,上述与上述目标用户会话以获取与上述目标用户相关的音频数据,可以包括以下步骤:
第一步,上述执行主体可以指示上述目标用户完成所要求的用户行为和/或回复所提出的目标问题。作为示例,上述所要求的用户行为可以是指示目标用户通过语音回复或者手动操作的方式来在目标应用中进行签到。上述所提出的目标问题可以是:“今天的心情如何?”。其中,上述目标应用可以是移动端应用(APP,Application)。
第二步,上述执行主体可以接收上述目标用户执行所要求的用户行为和/或回复所提出的目标问题而产生的音频数据。
可选地,上述与上述目标用户会话以获取与上述目标用户相关的音频数据,可以包括以下步骤:
第一步,指示上述目标用户对第二目标待展示信息进行语音反馈。其中,上述第二目标待展示信息相关的音频数据是预先设置的、随机插入至目标音频的音频数据。上述目标音频可以是目标应用所展示的音频。
第二步,接收上述目标用户相关的音频数据。
可选地,上述第二目标待展示信息相关的音频数据是通过以下步骤确定的。
第一步,响应于确定上述目标用户已登录上述目标应用,确定上述目标用户是否为第一次使用上述目标应用的用户。其中,上述目标应用可以是移动端应用(APP,Application)。
作为示例,响应于确定上述目标用户已登录上述目标应用,上述执行主体可以通过查询用户信息数据库来确定上述目标用户是否为第一次使用上述目标应用的用户。其中,上述用户信息数据库存储着目标用户的登录信息。
第二步,响应于确定不是,获取预先存储的与上述目标用户相关联的用户情绪画像。其中,上述用户情绪画像存储着目标用户历史情绪信息。
作为示例,上述执行主体可以从用户情绪画像数据库中获取预先存储的与上述目标用户相关联的用户情绪画像。
第三步,根据上述用户情绪画像,确定预先存储的各个待展示信息中的第二目标待展示信息。
作为示例,上述执行主体可以根据上述用户情绪画像,确定目标用户对应的,出现次数最多的某一类型情绪。然后,从上述各个待展示信息中确定出现次数最多的某一类型情绪最为匹配的第二目标待展示信息。
在这里,上述执行主体可以响应于接收到目标授权信号,获取上述用户情绪画像。其中,上述目标授权信号可以是上述用户画像对应的用户,对目标控件执行目标操作产生的信号。上述目标控件可以包含于授权提示框中。上述授权提示框可以在目标终端设备显示。上述目标终端设备可以是登录有上述用户对应账号的终端设备。上述终端设备可以是“手机”,也可以是“电脑”。上述目标操作可以是“点击操作”,也可以是“滑动操作”。上述目标控件可以是“确认按钮”。
作为示例,上述授权提示框可以如图3所示。上述授权提示框可以包括:提示信息显示部分301和控件302。其中,上述提示信息显 示部分301可以用于显示提示信息。上述提示信息可以是“是否允许获取用户情绪画像”。上述控件302可以是“确认按钮”,也可以是“取消按钮”。
第四步,匹配待播放的、与上述第二目标待展示信息相关的音频数据。
作为示例,上述执行主体可以通过预先设置的模板来匹配生成待播放的、与上述第二目标待展示信息相关的音频数据。
作为又一个示例,上述第二待展示信息可以是:“无限超脱,点滴创作”。对应的音频数据可以是:您是否需要了解与“无限超脱,点滴创作”相关的内容信息?。
步骤202,响应于确定上述音频数据不是噪声音频数据,生成上述目标用户的情绪信息。
在一些实施例中,响应于确定上述音频数据不是噪声音频数据,上述执行主体可以生成上述目标用户的情绪信息。其中,上述情绪信息可以包括但不限于以下至少一项:高兴,惊讶,悲伤,害怕,愤怒,厌恶。
作为示例,上述执行主体可以通过以下步骤来生成上述目标用户的情绪信息:
第一步,对上述音频数据进行降噪,可以得到降噪后的音频数据。
作为示例,上述执行主体可以利用噪声去除算法来对上述音频数据进行降噪,可以得到降噪后的音频数据。其中,上述噪声去除算法可以是最小值控制的递归平均算法(MCRA)。
第二步,利用自然语言处理方法,融合声学参数的统计特征和时序特征,对降噪后的音频数据进行分类,得到目标用户的情绪信息。
其中,声学参数可以通过以下步骤获取:
第一步,获取音频的韵律参数。其中,韵律参数包括:基频参数和时长参数。
作为示例,基频参数可以YIN算法来获得,时长参数可以通过相关标注工具来获得。
第二步,获取音频的频谱参数。其中,频谱参数包括:Mel频谱 倒谱参数MFCC(Mel Frequency Cepstrum Coefficient)、频谱质心参数Sc、频谱截止参数Sr、频谱变迁参数Sf、频带周期性参数Sp。其中,A(n)可以为第n条谱线所对应的幅度,Sc的计算公式可以是:
Sr的计算公式可以是:
记Ai(n)、Ai-1(n)分别为当前帧和前一帧的幅度谱,Sf的计算S公式可以是:
将语音信号通过若干个不同频率范围的滤波器,对于通过第j个频带的信号,在当前帧和前一帧范围内计算归一化相关函数,记Sj(m)为通过相应频带的信号,观察Rj(k)的平坦程度,Sp的计算公式可以是:
第三步,处理以上参数在音频数据中的统计特征,使用概率神经网络(PNN,Product Network)对统计特征进行识别。
第四步,处理以上参数在音频数据中的时序特征,使用隐马尔可夫模型(HMM,Hidden Markov Model)对时序特征进行识别。
第五步,从样本x中提取N组特征(包含统计和时序两种特征),记为f1~fN。通过PNN或HMM模型得到的属于第i类情感的概率为:P(C
i|f
1)~P(C
i|f
n),最终识别结果r可以是:
第六步,将F表示融合规则,按“乘法原则”和“加法原则”对两种特征的识别结果进行融合处理。根据已有研究,此算法在降低数据混淆度上表现较好。乘法规则公式可以是:
加法原则公式可以是:
在一些实施例的一些可选的实现方式中,上述步骤还包括:
将上述目标用户的情绪信息融入上述目标用户相关联的用户情绪画像。
在这里,上述目标用户相关联的用户情绪画像的积累,可以更为清楚用户生活中的情绪变化。以此,可以更有针对性的依据目标用户的用户情绪画像为目标用户提供待展示信息。
步骤203,确定与上述情绪信息相对应的待展示信息类型。
在一些实施例中,上述执行主体可以确定与上述情绪信息相对应的待展示信息类型。其中,待展示信息类型可以包括:待展示信息积极类型,待展示信息共情类型和待展示信息拒绝类型。上述待展示信息类型表征用户接收待展示信息的意愿程度。上述待展示信息积极类型可以表征用户较大意愿愿意接收待展示信息。上述待展示信息拒绝类型可以表征用户较小意愿愿意接收待展示信息。可以将不清楚用户是否愿意接收待展示信息的待展示信息类型确定为待展示信息共情类型。在这里,上述情绪信息与待展示信息类型可以是一一对应的关系。
作为示例,对于待展示信息积极类型,对应的情绪信息可以包括:高兴、惊讶、中性。对于待展示信息拒绝类型,对应的情绪信息可以包括:愤怒、厌恶。对于待展示信息共情类型,对应的情绪信息可以包括:悲伤、害怕。
作为示例,上述执行主体可以通过预先构建的情绪信息与待展示信息类型的关系表来确定与上述情绪信息相对应的待展示信息类型。
步骤204,响应于确定上述待展示信息类型不是目标待展示信息类型,确定待展示给上述目标用户的上述第一目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调。
在一些实施例中,响应于确定上述待展示信息类型不是目标待展示信息类型,上述执行主体可以确定待展示给上述目标用户的上述第一目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调。其中,上述目标待展示信息类型可以是待展示信息拒绝类型。因为当待展示信息类型为待展示信息拒绝类型,表征目标用户情绪信息可能为愤怒和厌恶。所以上述执行主体可以有选择性的推送第一目标待展示信息。上述待展示信息类型存在对应的第一目标待展示信息可以表征预先设置的、与待展示信息类型存在对应关系的待展示信息。作为示例,待展示信息积极类型对应着可以让用户保持愉悦情绪的待展示信息。例如,待展示信息可以是美妆护肤类的待展示信息。待展示信息拒绝类型对应着可以让用户消除消极情绪的待展示信息。例如,待展示信息可以是收纳用品类的待展示信息。待展示信息共情类型对应着让用户消除紧张情绪的待展示信息。例如,待展示信息可以是休闲食品类的待展示信息。
除此之外,待展示信息类型对应的提示语可以是预先设置的。例如,待展示信息积极类型对应的提示语可以是:“今天心情不错,推荐一本书给你哦!”。待展示信息共情类型对应的提示语可以是:“今天心情一般,来看看让人惊喜的小配饰吧~”。待展示信息拒绝类型对应提示语可以是:“谢谢”。
待展示信息类型对应的提示语的播放语调可以是预先设置的。例如,待展示信息积极类型对应的播放语调可以是积极、阳光的播放语调。待展示信息共情类型对应的提示语可以是平和、温柔的播放语调。待展示信息拒绝类型对应提示语可以是语气较为正式的播放语调。
步骤205,依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第一目标待展示信息展示给 上述目标用户。
在一些实施例中,上述执行主体可以依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第一目标待展示信息展示给上述目标用户。
本公开的上述各个实施例中具有如下有益效果:本公开的一些实施例的信息展示方法可以快捷、高效、更有针对性的将待展示信息推送给用户以供用户浏览,提高了用户体验。具体来说,不能高效、准确的确定出当前用户是否希望观看待展示信息。具体而言,在不确定用户是否希望观看待展示信息的情况下,展示待展示信息给目标用户,可能导致用户体验较差。基于此,本公开的一些实施例的信息展示方法可以首先获取与目标用户相关的音频数据以用于后续确定当前目标用户的情绪信息。然后,响应于确定上述音频数据不是噪声音频数据,生成上述目标用户的情绪信息。在这里,通过生成目标用户的情绪信息以此来进一步的确定当前用户接收待展示信息的意愿程度。进而,确定与上述情绪信息相对应的待展示信息类型。其中,上述待展示信息类型表征用户接收待展示信息的意愿程度。在这里,用户的情绪信息可以包括多种。直接通过用户的情绪信息来确定是否推送待展示信息和/或后续推送什么样的待展示信息可能比较繁琐,复杂。由此,通过确定与情绪信息相对应的待展示信息类型,可以更为具体、明确的确定出目标用户的意愿程度。接着,响应于确定上述待展示信息类型不是目标待展示信息类型,确定待展示给上述目标用户的、当前目标用户可能希望看到的第一目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调。可选地,目标待展示信息类型可以表征目标用户没有较大的意愿接收待展示信息。最后,依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第一目标待展示信息展示给上述目标用户。该信息展示方法充分考量了在展示待展示信息时,目标用户的情绪信息,使得目标用户观看到可能满足当前用户需求的待展示信息。以达到展示待展示信息的效果,侧面也极大程度的提高了用户体验。
参考图4,示出了根据本公开的信息展示方法的另一些实施例的流程400。该信息展示方法,包括以下步骤:
步骤401,获取与目标用户相关的音频数据。
步骤402,响应于确定上述音频数据不是噪声音频数据,生成上述目标用户的情绪信息。
步骤403,确定与上述情绪信息相对应的待展示信息类型。
步骤404,响应于确定上述待展示信息类型不是目标待展示信息类型,确定待展示给上述目标用户的第一目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调。
步骤405,依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第一目标待展示信息展示给上述目标用户。
在一些实施例中,步骤401-405的具体实现及其所带来的技术效果,可以参考图2对应的实施例中的步骤201-205,在此不再赘述。
步骤406,响应于确定上述音频数据为上述噪声音频数据,获取上述目标用户的用户行为信息。
在一些实施例中,响应于确定上述音频数据为上述噪声音频数据,执行主体(例如图1所示的电子设备)可以通过查询用户信息数据库来获取上述目标用户的用户行为信息。其中,上述用户行为信息可以包括但不限于以下至少一项:用户的页面点击信息,用户的页面浏览信息,用户执行第一价值转移操作(购买)的信息、用户执行第二价值转移操作(转账)的信息。
作为示例,上述执行主体可以通过相关自然语言处理算法来确定音频数据是否为上述噪声音频数据。
步骤407,确定与上述用户行为信息相对应的待展示信息类型。
在一些实施例中,上述执行主体可以确定与上述用户行为信息相对应的待展示信息类型。
作为示例,上述执行主体可以通过预先建立的用户行为信息与待展示信息类型之间关联关系的表来确定与上述用户行为信息相对应的待展示信息类型。
作为示例,上述执行主体可以通过统计各种用户行为信息出现的次数。然后,通过上述表来确定各个用户行为信息对应各个待展示信息类型出现的次数。最后,将各个待展示信息类型出现的次数最多的待展示信息类型确定为与上述用户行为信息相对应的待展示信息类型。
步骤408,响应于确定上述待展示信息类型不是上述目标待展示信息类型,确定待展示给上述目标用户的、与上述待展示信息类型相关联的第三目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调。
步骤409,依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第三目标待展示信息展示给上述目标用户。
在一些实施例中,步骤408-409的具体实现及其所带来的技术效果,可以参考图2对应的实施例中的步骤204-205,在此不再赘述。
从图4中可以看出,与图2对应的一些实施例的描述相比,图4对应的一些实施例中的信息展示方法的流程400更加突出了根据用户行为信息来展示待展示信息的具体步骤。由此,这些实施例描述的方案可以在不需要实时获取有效的用户音频数据的情况下,可以通过分析用户行为信息,更有针对性、高效的推送待展示信息,侧面提高了用户体验。
参考图5,作为对上述各图所示方法的实现,本公开提供了一种信息展示装置的一些实施例,这些装置实施例与图2所示的那些方法实施例相对应,该装置具体可以应用于各种电子设备中。
如图5所示,一种信息展示装置500包括:获取单元501、生成单元502、第一确定单元503、第二确定单元504和展示单元505。其中,获取单元501,被配置成获取与目标用户相关的音频数据;生成单元502,被配置成响应于确定上述音频数据不是噪声音频数据,生成上述目标用户的情绪信息;第一确定单元503,被配置成确定与上述情绪信息相对应的待展示信息类型,其中,上述待展示信息类型表 征用户接收待展示信息的意愿程度;第二确定单元504,被配置成响应于确定上述待展示信息类型不是目标待展示信息类型,确定待展示给上述目标用户的第一目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调;展示单元505,被配置成依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第一目标待展示信息展示给上述目标用户。
在一些实施例的一些可选的实现方式中,上述获取单元501可以被配置成:与上述目标用户会话以获取与上述目标用户相关的音频数据。
在一些实施例的一些可选的实现方式中,上述获取单元501可以被配置成:指示上述目标用户完成所要求的用户行为和/或回复所提出的目标问题;接收上述目标用户执行所要求的用户行为和/或回复所提出的目标问题而产生的音频数据。
在一些实施例的一些可选的实现方式中,上述获取单元501可以被配置成:指示上述目标用户对上述第二目标待展示信息进行语音反馈,其中,上述第二目标待展示信息相关的音频数据是预先设置的、随机插入至目标音频的音频数据;接收上述目标用户相关的音频数据。
在一些实施例的一些可选的实现方式中,上述装置500还包括:用户信息获取单元、第三确定单元、第四确定单元、播放展示单元(图中未显示)。其中,用户信息获取单元可以被配置成:响应于确定上述音频数据为上述噪声音频数据,获取上述目标用户的用户行为信息。第三确定单元可以被配置成:确定与上述用户行为信息相对应的待展示信息类型。第四确定单元可以被配置成:响应于确定上述待展示信息类型不是上述目标待展示信息类型,确定待展示给上述目标用户的、与上述待展示信息类型相关联的第三目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调。播放展示单元可以被配置成:依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第三目标待展示信息展示给上述目标用户。
在一些实施例的一些可选的实现方式中,上述装置500还包括: 播放单元(图中未显示)。其中,播放单元可以被配置成:响应于确定上述待展示信息类型是上述目标待展示信息类型,使用目标语调来播放结束语。
在一些实施例的一些可选的实现方式中,上述第二目标待展示信息相关的音频数据是通过以下步骤确定的:响应于确定上述目标用户已登录上述目标应用,确定上述目标用户是否为第一次使用上述目标应用的用户;响应于确定不是,获取预先存储的与上述目标用户相关联的用户情绪画像;根据上述用户情绪画像,确定预先存储的各个待展示信息中的第二目标待展示信息;匹配待播放的、与上述第二目标待展示信息相关的音频数据。
在一些实施例的一些可选的实现方式中,上述装置500还包括:融入单元(图中未显示)。其中,融入单元可以被配置成:将上述目标用户的情绪信息融入上述目标用户相关联的用户情绪画像。
可以理解的是,该装置500中记载的诸单元与参考图2描述的方法中的各个步骤相对应。由此,上文针对方法描述的操作、特征以及产生的有益效果同样适用于装置500及其中包含的单元,在此不再赘述。
参考图6,其示出了适于用来实现本公开的一些实施例的电子设备(例如图1中的电子设备)600的结构示意图。图6示出的电子设备仅仅是一个示例,不应对本公开的实施例的功能和使用范围带来任何限制。
如图6所示,电子设备600可以包括处理装置(例如中央处理器、图形处理器等)601,其可以根据存储在只读存储器(ROM)602中的程序或者从存储装置608加载到随机访问存储器(RAM)603中的程序而执行各种适当的动作和处理。在RAM 603中,还存储有电子设备600操作所需的各种程序和数据。处理装置601、ROM 602以及RAM603通过总线604彼此相连。输入/输出(I/O)接口605也连接至总线604。
通常,以下装置可以连接至I/O接口605:包括例如触摸屏、触摸 板、键盘、鼠标、摄像头、麦克风、加速度计、陀螺仪等的输入装置606;包括例如液晶显示器(LCD)、扬声器、振动器等的输出装置607;包括例如磁带、硬盘等的存储装置608;以及通信装置609。通信装置609可以允许电子设备600与其他设备进行无线或有线通信以交换数据。虽然图6示出了具有各种装置的电子设备600,但是应理解的是,并不要求实施或具备所有示出的装置。可以替代地实施或具备更多或更少的装置。图6中示出的每个方框可以代表一个装置,也可以根据需要代表多个装置。
特别地,根据本公开的一些实施例,上文参考流程图描述的过程可以被实现为计算机软件程序。例如,本公开的一些实施例包括一种计算机程序产品,其包括承载在计算机可读介质上的计算机程序,该计算机程序包含用于执行流程图所示的方法的程序代码。在这样的一些实施例中,该计算机程序可以通过通信装置609从网络上被下载和安装,或者从存储装置608被安装,或者从ROM 602被安装。在该计算机程序被处理装置601执行时,执行本公开的一些实施例的方法中限定的上述功能。
需要说明的是,本公开的一些实施例上述的计算机可读介质可以是计算机可读信号介质或者计算机可读存储介质或者是上述两者的任意组合。计算机可读存储介质例如可以是——但不限于——电、磁、光、电磁、红外线、或半导体的系统、装置或器件,或者任意以上的组合。计算机可读存储介质的更具体的例子可以包括但不限于:具有至少一个导线的电连接、便携式计算机磁盘、硬盘、随机访问存储器(RAM)、只读存储器(ROM)、可擦式可编程只读存储器(EPROM或闪存)、光纤、便携式紧凑磁盘只读存储器(CD-ROM)、光存储器件、磁存储器件、或者上述的任意合适的组合。在本公开的一些实施例中,计算机可读存储介质可以是任何包含或存储程序的有形介质,该程序可以被指令执行系统、装置或者器件使用或者与其结合使用。而在本公开的一些实施例中,计算机可读信号介质可以包括在基带中或者作为载波一部分传播的数据信号,其中承载了计算机可读的程序代码。这种传播的数据信号可以采用多种形式,包括但不限于电磁信 号、光信号或上述的任意合适的组合。计算机可读信号介质还可以是计算机可读存储介质以外的任何计算机可读介质,该计算机可读信号介质可以发送、传播或者传输用于由指令执行系统、装置或者器件使用或者与其结合使用的程序。计算机可读介质上包含的程序代码可以用任何适当的介质传输,包括但不限于:电线、光缆、RF(射频)等等,或者上述的任意合适的组合。
在一些实施方式中,客户端、服务器可以利用诸如HTTP(HyperText Transfer Protocol,超文本传输协议)之类的任何当前已知或未来研发的网络协议进行通信,并且可以与任意形式或介质的数字数据通信(例如,通信网络)互连。通信网络的示例包括局域网(“LAN”),广域网(“WAN”),网际网(例如,互联网)以及端对端网络(例如,ad hoc端对端网络),以及任何当前已知或未来研发的网络。
上述计算机可读介质可以是上述电子设备中所包含的;也可以是单独存在,而未装配入该电子设备中。上述计算机可读介质承载有至少一个程序,当上述至少一个程序被该电子设备执行时,使得该电子设备:获取与目标用户相关的音频数据;响应于确定上述音频数据不是噪声音频数据,生成上述目标用户的情绪信息;确定与上述情绪信息相对应的待展示信息类型,其中,上述待展示信息类型表征用户接收待展示信息的意愿程度;响应于确定上述待展示信息类型不是目标待展示信息类型,确定待展示给上述目标用户的第一目标待展示信息、上述待展示信息类型对应的提示语和上述提示语的播放语调;依照上述播放语调,播放上述提示语给上述目标用户,以及响应于上述提示语播放结束,将上述第一目标待展示信息展示给上述目标用户。
可以以一种或多种程序设计语言或其组合来编写用于执行本公开的一些实施例的操作的计算机程序代码,上述程序设计语言包括面向对象的程序设计语言—诸如Java、Smalltalk、C++,还包括常规的过程式程序设计语言—诸如“C”语言或类似的程序设计语言。程序代码可以完全地在用户计算机上执行、部分地在用户计算机上执行、作为一个独立的软件包执行、部分在用户计算机上部分在远程计算机上执行、 或者完全在远程计算机或服务器上执行。在涉及远程计算机的情形中,远程计算机可以通过任意种类的网络——包括局域网(LAN)或广域网(WAN)——连接到用户计算机,或者,可以连接到外部计算机(例如利用因特网服务提供商来通过因特网连接)。
附图中的流程图和框图,图示了按照本公开各种实施例的系统、方法和计算机程序产品的可能实现的体系架构、功能和操作。在这点上,流程图或框图中的每个方框可以代表一个模块、程序段、或代码的一部分,该模块、程序段、或代码的一部分包含至少一个用于实现规定的逻辑功能的可执行指令。也应当注意,在有些作为替换的实现中,方框中所标注的功能也可以以不同于附图中所标注的顺序发生。例如,两个接连地表示的方框实际上可以基本并行地执行,它们有时也可以按相反的顺序执行,这依所涉及的功能而定。也要注意的是,框图和/或流程图中的每个方框、以及框图和/或流程图中的方框的组合,可以用执行规定的功能或操作的专用的基于硬件的系统来实现,或者可以用专用硬件与计算机指令的组合来实现。
描述于本公开的一些实施例中的单元可以通过软件的方式实现,也可以通过硬件的方式来实现。所描述的单元也可以设置在处理器中,例如,可以描述为:一种处理器包括获取单元、生成单元、第一确定单元、第二确定单元和展示单元。其中,这些单元的名称在某种情况下并不构成对该单元本身的限定,例如,获取单元还可以被描述为“获取与目标用户相关的音频数据的单元”。
本文中以上描述的功能可以至少部分地由至少一个硬件逻辑部件来执行。例如,非限制性地,可以使用的示范类型的硬件逻辑部件包括:现场可编程门阵列(FPGA)、专用集成电路(ASIC)、专用标准产品(ASSP)、片上系统(SOC)、复杂可编程逻辑设备(CPLD)等等。
以上描述仅为本公开的一些较佳实施例以及对所运用技术原理的说明。本领域技术人员应当理解,本公开的实施例中所涉及的发明范围,并不限于上述技术特征的特定组合而成的技术方案,同时也应涵盖在不脱离上述发明构思的情况下,由上述技术特征或其等同特征进 行任意组合而形成的其它技术方案。例如上述特征与本公开的实施例中公开的(但不限于)具有类似功能的技术特征进行互相替换而形成的技术方案。
Claims (11)
- 一种信息展示方法,包括:获取与目标用户相关的音频数据;响应于确定所述音频数据不是噪声音频数据,生成所述目标用户的情绪信息;确定与所述情绪信息相对应的待展示信息类型,其中,所述待展示信息类型表征用户接收待展示信息的意愿程度;响应于确定所述待展示信息类型不是目标待展示信息类型,确定待展示给所述目标用户的第一目标待展示信息、所述待展示信息类型对应的提示语和所述提示语的播放语调;依照所述播放语调,播放所述提示语给所述目标用户,以及响应于所述提示语播放结束,将所述第一目标待展示信息展示给所述目标用户。
- 根据权利要求1所述的方法,其中,所述获取与目标用户相关的音频数据,包括:与所述目标用户会话以获取与所述目标用户相关的音频数据。
- 根据权利要求2所述的方法,其中,所述与所述目标用户会话以获取与所述目标用户相关的音频数据,包括:指示所述目标用户完成所要求的用户行为和/或回复所提出的目标问题;接收所述目标用户执行所要求的用户行为和/或回复所提出的目标问题而产生的音频数据。
- 根据权利要求2或3所述的方法,其中,所述与所述目标用户会话以获取与所述目标用户相关的音频数据,包括:指示所述目标用户对第二目标待展示信息进行语音反馈,其中,所述第二目标待展示信息相关的音频数据是预先设置的、随机插入至 目标音频的音频数据;接收所述目标用户相关的音频数据。
- 根据权利要求1-4之一所述的方法,其中,所述方法还包括:响应于确定所述音频数据为所述噪声音频数据,获取所述目标用户的用户行为信息;确定与所述用户行为信息相对应的待展示信息类型;响应于确定所述待展示信息类型不是所述目标待展示信息类型,确定待展示给所述目标用户的、与所述待展示信息类型相关联的第三目标待展示信息、所述待展示信息类型对应的提示语和所述提示语的播放语调;依照所述播放语调,播放所述提示语给所述目标用户,以及响应于所述提示语播放结束,将所述第三目标待展示信息展示给所述目标用户。
- 根据权利要求1-5之一所述的方法,其中,所述方法还包括:响应于确定所述待展示信息类型是所述目标待展示信息类型,使用目标语调来播放结束语。
- 根据权利要求4所述的方法,其中,所述第二目标待展示信息相关的音频数据是通过以下步骤确定的:响应于确定所述目标用户已登录所述目标应用,确定所述目标用户是否为第一次使用所述目标应用的用户;响应于确定不是,获取预先存储的与所述目标用户相关联的用户情绪画像;根据所述用户情绪画像,确定预先存储的各个待展示信息中的第二目标待展示信息;匹配待播放的、与所述第二目标待展示信息相关的音频数据。
- 根据权利要求1-7之一所述的方法,其中,在响应于确定所述 音频数据不是噪声音频数据,生成所述目标用户的情绪信息之后,所述方法还包括:将所述目标用户的情绪信息融入所述目标用户相关联的用户情绪画像。
- 一种信息展示装置,包括:获取单元,被配置成获取与目标用户相关的音频数据;生成单元,被配置成响应于确定所述音频数据不是噪声音频数据,生成所述目标用户的情绪信息;第一确定单元,被配置成确定与所述情绪信息相对应的待展示信息类型,其中,所述待展示信息类型表征用户接收待展示信息的意愿程度;第二确定单元,被配置成响应于确定所述待展示信息类型不是目标待展示信息类型,确定待展示给所述目标用户的第一目标待展示信息、所述待展示信息类型对应的提示语和所述提示语的播放语调;展示单元,被配置成依照所述播放语调,播放所述提示语给所述目标用户,以及响应于所述提示语播放结束,将所述第一目标待展示信息展示给所述目标用户。
- 一种电子设备,包括:至少一个处理器;存储装置,其上存储有至少一个程序,当所述至少一个程序被所述至少一个处理器执行,使得所述至少一个处理器实现如权利要求1-8中任一所述的方法。
- 一种计算机可读介质,其上存储有计算机程序,其中,所述程序被处理器执行时实现如权利要求1-8中任一所述的方法。
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| CN202110426791.4 | 2021-04-20 | ||
| CN202110426791.4A CN115312079A (zh) | 2021-04-20 | 2021-04-20 | 信息展示方法、装置、电子设备和计算机可读介质 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2022222841A1 true WO2022222841A1 (zh) | 2022-10-27 |
Family
ID=83721954
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/CN2022/086827 Ceased WO2022222841A1 (zh) | 2021-04-20 | 2022-04-14 | 信息展示方法、装置、电子设备和计算机可读介质 |
Country Status (2)
| Country | Link |
|---|---|
| CN (1) | CN115312079A (zh) |
| WO (1) | WO2022222841A1 (zh) |
Cited By (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN116932919A (zh) * | 2023-09-15 | 2023-10-24 | 中关村科学城城市大脑股份有限公司 | 信息推送方法、装置、电子设备和计算机可读介质 |
Families Citing this family (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN117115532B (zh) * | 2023-08-23 | 2024-01-26 | 广州一线展示设计有限公司 | 一种基于物联网的展览台智能控制方法及系统 |
Citations (8)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2003271194A (ja) * | 2002-03-14 | 2003-09-25 | Canon Inc | 音声対話装置及びその制御方法 |
| CN105654950A (zh) * | 2016-01-28 | 2016-06-08 | 百度在线网络技术(北京)有限公司 | 自适应语音反馈方法和装置 |
| CN108154398A (zh) * | 2017-12-27 | 2018-06-12 | 广东欧珀移动通信有限公司 | 信息显示方法、装置、终端及存储介质 |
| CN108648768A (zh) * | 2018-04-16 | 2018-10-12 | 广州市菲玛尔咨询服务有限公司 | 一种咨询推荐方法及其管理系统 |
| CN109346076A (zh) * | 2018-10-25 | 2019-02-15 | 三星电子(中国)研发中心 | 语音交互、语音处理方法、装置和系统 |
| CN109949071A (zh) * | 2019-01-31 | 2019-06-28 | 平安科技(深圳)有限公司 | 基于语音情绪分析的产品推荐方法、装置、设备和介质 |
| CN109961786A (zh) * | 2019-01-31 | 2019-07-02 | 平安科技(深圳)有限公司 | 基于语音分析的产品推荐方法、装置、设备和存储介质 |
| CN110188177A (zh) * | 2019-05-28 | 2019-08-30 | 北京搜狗科技发展有限公司 | 对话生成方法及装置 |
Family Cites Families (7)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN106682063B (zh) * | 2016-10-20 | 2018-04-03 | 北京跃盟科技有限公司 | 一种广告信息推送方法、装置以及系统 |
| CN109065035A (zh) * | 2018-09-06 | 2018-12-21 | 珠海格力电器股份有限公司 | 信息交互方法及装置 |
| CN111222044A (zh) * | 2019-12-31 | 2020-06-02 | 深圳Tcl数字技术有限公司 | 基于情绪感知的信息推荐方法、设备及存储介质 |
| CN111883131B (zh) * | 2020-08-20 | 2023-10-27 | 腾讯科技(深圳)有限公司 | 语音数据的处理方法及装置 |
| CN112185389B (zh) * | 2020-09-22 | 2024-06-18 | 北京小米松果电子有限公司 | 语音生成方法、装置、存储介质和电子设备 |
| CN112182173B (zh) * | 2020-09-23 | 2024-08-06 | 支付宝(杭州)信息技术有限公司 | 一种基于虚拟生命的人机交互方法、装置及电子设备 |
| CN112667196B (zh) * | 2021-01-28 | 2024-08-13 | 百度在线网络技术(北京)有限公司 | 信息展示方法及装置、电子设备和介质 |
-
2021
- 2021-04-20 CN CN202110426791.4A patent/CN115312079A/zh active Pending
-
2022
- 2022-04-14 WO PCT/CN2022/086827 patent/WO2022222841A1/zh not_active Ceased
Patent Citations (8)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| JP2003271194A (ja) * | 2002-03-14 | 2003-09-25 | Canon Inc | 音声対話装置及びその制御方法 |
| CN105654950A (zh) * | 2016-01-28 | 2016-06-08 | 百度在线网络技术(北京)有限公司 | 自适应语音反馈方法和装置 |
| CN108154398A (zh) * | 2017-12-27 | 2018-06-12 | 广东欧珀移动通信有限公司 | 信息显示方法、装置、终端及存储介质 |
| CN108648768A (zh) * | 2018-04-16 | 2018-10-12 | 广州市菲玛尔咨询服务有限公司 | 一种咨询推荐方法及其管理系统 |
| CN109346076A (zh) * | 2018-10-25 | 2019-02-15 | 三星电子(中国)研发中心 | 语音交互、语音处理方法、装置和系统 |
| CN109949071A (zh) * | 2019-01-31 | 2019-06-28 | 平安科技(深圳)有限公司 | 基于语音情绪分析的产品推荐方法、装置、设备和介质 |
| CN109961786A (zh) * | 2019-01-31 | 2019-07-02 | 平安科技(深圳)有限公司 | 基于语音分析的产品推荐方法、装置、设备和存储介质 |
| CN110188177A (zh) * | 2019-05-28 | 2019-08-30 | 北京搜狗科技发展有限公司 | 对话生成方法及装置 |
Cited By (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN116932919A (zh) * | 2023-09-15 | 2023-10-24 | 中关村科学城城市大脑股份有限公司 | 信息推送方法、装置、电子设备和计算机可读介质 |
| CN116932919B (zh) * | 2023-09-15 | 2023-11-24 | 中关村科学城城市大脑股份有限公司 | 信息推送方法、装置、电子设备和计算机可读介质 |
Also Published As
| Publication number | Publication date |
|---|---|
| CN115312079A (zh) | 2022-11-08 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US11475897B2 (en) | Method and apparatus for response using voice matching user category | |
| US10623573B2 (en) | Personalized support routing based on paralinguistic information | |
| CN116127046A (zh) | 生成式大语言模型训练方法、基于模型的人机语音交互方法 | |
| WO2022121257A1 (zh) | 模型训练方法、语音识别方法、装置、设备及存储介质 | |
| CN111798821B (zh) | 声音转换方法、装置、可读存储介质及电子设备 | |
| CN111801645A (zh) | 具有经改进交互式动画会话界面系统的计算装置 | |
| CN107818798A (zh) | 客服服务质量评价方法、装置、设备及存储介质 | |
| US12265756B2 (en) | Graphical interface for speech-enabled processing | |
| CN111785247A (zh) | 语音生成方法、装置、设备和计算机可读介质 | |
| CN109256133A (zh) | 一种语音交互方法、装置、设备及存储介质 | |
| CN112002346A (zh) | 基于语音的性别年龄识别方法、装置、设备和存储介质 | |
| CN108922525A (zh) | 语音处理方法、装置、存储介质及电子设备 | |
| CN115222857A (zh) | 生成虚拟形象的方法、装置、电子设备和计算机可读介质 | |
| US20250232768A1 (en) | System method and apparatus for combining words and behaviors | |
| CN112364144A (zh) | 交互方法、装置、设备和计算机可读介质 | |
| WO2022222841A1 (zh) | 信息展示方法、装置、电子设备和计算机可读介质 | |
| US20250078574A1 (en) | Automatic sign language interpreting | |
| CN116741143B (zh) | 基于数字分身的个性化ai名片的交互方法及相关组件 | |
| CN118520940A (zh) | 知识图谱构建方法、装置、存储介质以及终端 | |
| US20220130413A1 (en) | Systems and methods for a computerized interactive voice companion | |
| Lahiri et al. | Hybrid multi purpose voice assistant | |
| CN119646127A (zh) | 智能问答方法、装置、电子设备及存储介质 | |
| EP3846164B1 (en) | Method and apparatus for processing voice, electronic device, storage medium, and computer program product | |
| EP4139784B1 (en) | Hierarchical context specific actions from ambient speech | |
| CN110086937A (zh) | 通话界面的显示方法、电子设备和计算机可读介质 |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 22790949 Country of ref document: EP Kind code of ref document: A1 |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 32PN | Ep: public notification in the ep bulletin as address of the adressee cannot be established |
Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205A DATED 15-02-2024) |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 22790949 Country of ref document: EP Kind code of ref document: A1 |





