WO2024178590A1 - 一种生成虚拟形象的方法和装置 - Google Patents
一种生成虚拟形象的方法和装置 Download PDFInfo
- Publication number
- WO2024178590A1 WO2024178590A1 PCT/CN2023/078643 CN2023078643W WO2024178590A1 WO 2024178590 A1 WO2024178590 A1 WO 2024178590A1 CN 2023078643 W CN2023078643 W CN 2023078643W WO 2024178590 A1 WO2024178590 A1 WO 2024178590A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- information
- user
- virtual image
- text content
- characteristic information
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06T—IMAGE DATA PROCESSING OR GENERATION, IN GENERAL
- G06T13/00—Animation
- G06T13/80—Two-dimensional [2D] animation, e.g. using sprites
Definitions
- Embodiments of the present application relate to the field of human-computer interaction, and more specifically, to a method and device for generating a virtual image.
- Virtual images are becoming more and more popular because they can provide users with diversified styles and intelligent interactive services.
- the terminal device can provide users with a variety of style parameters for users to customize. Since there are many style parameters and there is no correlation between the parameters, users need to have a certain knowledge base on the setting of virtual images in order to set up a virtual image that meets their expectations. This will increase the user's learning cost and affect the user's setting efficiency and usage experience.
- the present application provides a method and device for generating a virtual image, which helps to reduce the learning cost of users when setting up a virtual image, improves the efficiency of setting up the virtual image, and also helps to improve the user experience.
- the method provided in the present application can be applied to mobile phones, tablet computers, wearable devices, augmented reality (AR)/virtual reality (VR) devices, laptops, ultra-mobile personal computers (UMPC), netbooks, personal digital assistants (PDA), vehicles and other terminal devices.
- AR augmented reality
- VR virtual reality
- laptops laptops
- ultra-mobile personal computers UMPC
- netbooks netbooks
- PDA personal digital assistants
- the embodiments of the present application do not impose any restrictions on the specific types of terminal devices.
- the vehicle is a vehicle in a broad sense, which can be a means of transportation (such as commercial vehicles, passenger cars, motorcycles, flying cars, trains, etc.), industrial vehicles (such as forklifts, trailers, tractors, etc.), engineering vehicles (such as excavators, bulldozers, cranes, etc.), agricultural equipment (such as mowers, harvesters, etc.), amusement equipment, toy vehicles, etc.
- transportation such as commercial vehicles, passenger cars, motorcycles, flying cars, trains, etc.
- industrial vehicles such as forklifts, trailers, tractors, etc.
- engineering vehicles such as excavators, bulldozers, cranes, etc.
- agricultural equipment such as mowers, harvesters, etc.
- amusement equipment toy vehicles, etc.
- toy vehicles etc.
- the embodiments of the present application do not specifically limit the types of vehicles.
- a method for generating a virtual image comprising: obtaining first text content; determining first personality characteristic information based on the first text content and a user's historical information record; and generating a virtual image corresponding to the first text content based on the first personality characteristic information.
- the terminal device can generate a virtual image based on the text content input by the user and the user's historical information record.
- a virtual image that better meets the user's expectations can be generated in combination with the user's historical information record, which also makes it easier for the user to resonate with the virtual image.
- the user does not need to choose from a large number of unrelated style parameters, which helps to reduce the user's learning cost when setting the virtual image, improves the efficiency of setting the virtual image, and also helps to improve the user's experience.
- the first text content includes a person's name or nickname.
- the user's historical information records include the user's search records for news in news applications, the use of efficiency applications (such as work plan applications, speech-to-text applications, etc.), the search records for movies or TV series in video applications, the launch of games, etc., in the past period of time (for example, one month).
- efficiency applications such as work plan applications, speech-to-text applications, etc.
- the search records for movies or TV series in video applications the launch of games, etc., in the past period of time (for example, one month).
- the user's historical information record also includes the duration and frequency of use of different applications. For example, a user who frequently uses news applications may prefer a serious format text; a user who frequently uses video applications may prefer a novel format text; a user who frequently uses efficiency applications may prefer a concise and brief format text.
- a correspondence between personality characteristic information and a virtual image is stored in a terminal device, and a virtual image corresponding to the first text content is generated according to the first personality characteristic information, including: generating the virtual image according to the first personality characteristic and the correspondence.
- determining the first personality characteristic information based on the first text content and the user's historical information record includes: when the first text content is included in the historical information record, determining, based on the first text content, text co-occurrence words corresponding to the first text content; and determining the first personality characteristic information based on the text co-occurrence words.
- the first personality feature information can be determined by the text co-occurrence words corresponding to the first text content.
- the personality feature information of the virtual image can be more in line with the user's expectations, and the final generated virtual image can be more in line with the user's expectations, which helps to improve the user's experience.
- determining the first personality characteristic information based on the co-occurring words in the text includes: determining the first personality characteristic information corresponding to the first text content based on text similarities between the co-occurring words in the text and words related to different personality dimensions.
- determining the first personality characteristic information based on the first text content and the user's historical information record includes: when the first text content is not included in the historical information record, determining the similarity between the first text content and each of one or more names in the historical information record; and determining the first personality characteristic information based on the similarity of each name and the personality characteristic information corresponding to each name.
- the personality feature information corresponding to the first text content can be obtained by calculating the similarity with each name in the historical information record.
- the personality feature information of the virtual image can be made more in line with the user's expectations, and the final generated virtual image can also be made more in line with the user's expectations, which helps to improve the user's experience.
- the method further includes: when it is detected that the user issues a voice command or a text input command, controlling the virtual image to make a voice reply or a text reply through a second text content, and the second text content is determined by the historical information record.
- generating a virtual image corresponding to the first text content based on the first personality characteristic information includes: prompting a user with multiple personality characteristic information based on the first text content and the historical information record; and determining the first personality characteristic information from the multiple personality characteristic information based on user input of the multiple personality characteristic information.
- generating a virtual image corresponding to the first text content according to the first personality characteristic information includes: determining first style characteristic information according to the first personality characteristic information; and generating the virtual image according to the first style characteristic information.
- a correspondence between personality feature information and style feature information is stored in a terminal device, and determining the first style feature information according to the first personality feature information includes: determining the first style feature information according to the first personality feature information and the correspondence.
- the style feature information required for generating a virtual image can be quickly obtained and the corresponding virtual image can be generated.
- the user does not need to select from a large number of unrelated style parameters, which helps to reduce the learning cost of the user when setting the virtual image, improves the efficiency of setting the virtual image, and also helps to improve the user's experience.
- determining first style feature information based on the first personality feature information includes: inputting the first personality feature information into a prediction model to obtain the first style feature information; wherein the prediction model is trained by sample training data, the sample training data includes sample names, sample personality feature information, and sample style feature information, and the sample style feature information includes at least one of an avatar, a body shape, a timbre, a speaking speed, a tone, and an expression corresponding to the sample name.
- the prediction model is trained through sample names, sample personality feature information and sample style feature information corresponding to the sample names, so that the prediction model can include different style feature dimensions with names as alignment labels, which helps to ensure the personality consistency of various style parameters of the virtual image, and can also make the final virtual image more in line with the user's expectations, which helps to improve the user experience.
- a virtual image corresponding to the first text content is generated based on the first style feature information, including: prompting the user with style feature information of one or more dimensions; and generating the virtual image based on the user's input of the style feature information of the one or more dimensions.
- the user can participate in the generation process of the virtual image, and the virtual image can be more in line with the user's expectations, which helps to improve the user's usage experience.
- determining the first style feature information according to the first personality feature information includes: determining the style feature information of the one or more dimensions according to the first personality feature information.
- the style feature information of one or more dimensions may be a style parameter list of one or more dimensions.
- the one or more dimensions of style feature information include at least one of avatar information, shape information, timbre information, speaking speed information, intonation information, and expression information.
- the avatar information may be an avatar list consisting of multiple avatars
- the body information may be a body list consisting of multiple bodies
- the timbre information may be a timbre list consisting of multiple timbres
- the speaking rate information may be a speaking rate list consisting of multiple speaking rates
- the tone information may be a tone list consisting of multiple tones
- the expression information may be an expression list consisting of multiple expressions.
- generating a virtual image corresponding to the first text content according to the first personality characteristic information includes: prompting a user with multiple virtual images according to the first personality characteristic information; According to the user's input to the multiple virtual images, a virtual image corresponding to the first text content is determined from the multiple virtual images.
- the user can participate in the generation process of the virtual image, and the virtual image can be more in line with the user's expectations, which helps to improve the user's usage experience.
- a device for generating a virtual image comprising: an acquisition unit for acquiring a first text content; a determination unit for determining first personality characteristic information based on the first text content and a user's historical information record; and a virtual image generation unit for generating a virtual image corresponding to the first text content based on the first personality characteristic information.
- the determination unit is used to: when the first text content is included in the historical information record, determine, based on the first text content, text co-occurrence words corresponding to the first text content; and determine the first personality characteristic information based on the text co-occurrence words.
- the determination unit is used to: when the first text content is not included in the historical information record, determine the similarity between the first text content and each of the one or more names in the historical information record; and determine the first personality characteristic information based on the similarity of each name and the personality characteristic information corresponding to each name.
- the device also includes a detection unit and a control unit, the detection unit being used to detect that a user issues a voice command or a text input command; the control unit being used to control the virtual image to make a voice reply or a text reply through a second text content, and the second text content is determined by the historical information record.
- the device also includes: a first prompting unit, used to prompt a user with multiple personality characteristic information based on the first text content and the historical information record; and the determining unit, used to determine the first personality characteristic information from the multiple personality characteristic information based on the user's input of the multiple personality characteristic information.
- the determination unit is further used to determine first style feature information based on the first personality feature information; and the virtual image generation unit is used to generate the virtual image based on the first style feature information.
- the determination unit is used to: input the first personality characteristic information into a prediction model to obtain the first style characteristic information; wherein the prediction model is trained by sample training data, the sample training data includes sample names, sample personality characteristic information, and sample style characteristic information, and the sample style characteristic information includes at least one of an avatar, shape, timbre, speaking speed, intonation, and expression corresponding to the sample name.
- the device also includes: a second prompting unit, used to prompt the user with style feature information of one or more dimensions; and the virtual image generation unit, used to generate the virtual image based on the user's input of the style feature information of the one or more dimensions.
- the one or more dimensions of style feature information include at least one of avatar information, shape information, timbre information, speaking speed information, intonation information, and expression information.
- the device also includes: a third prompting unit, used to prompt a plurality of virtual images to the user based on the first personality characteristic information; and the virtual image generating unit, used to determine the virtual image corresponding to the first text content from the plurality of virtual images based on the user's input to the plurality of virtual images.
- the present application provides a device for generating a virtual image, the device comprising a processing unit and a storage unit, wherein the storage unit is used to store instructions, and the processing unit executes the instructions stored in the storage unit so that the device executes any possible method in the first aspect.
- the present application provides a terminal device, which includes any possible device in the second aspect, or includes the device described in the third aspect.
- the terminal device is a vehicle, a computer, or a mobile phone.
- the present application provides a computer program product, comprising: a computer program code, when the computer program code is run on a computer, the computer executes any possible method in the first aspect above.
- the above-mentioned computer program code can be stored in whole or in part on the first storage medium, wherein the first storage medium can be packaged together with the processor or separately packaged with the processor, and the embodiments of the present application do not specifically limit this.
- the present application provides a computer-readable medium storing a program code, and when the computer program code runs on a computer, the computer executes any possible method in the first aspect.
- the present application provides a chip, comprising a circuit, wherein the circuit is used to execute any possible method in the first aspect above.
- FIG1 is a functional block diagram of a terminal device provided in an embodiment of the present application.
- FIG. 2 is a schematic diagram of the distribution of display screens in a vehicle cabin provided in an embodiment of the present application.
- FIG3 is a schematic flowchart of a method for generating a virtual image provided in an embodiment of the present application.
- FIG. 4 is a graphical user interface GUI provided in an embodiment of the present application.
- FIG. 5 is another set of GUIs provided by an embodiment of the present application.
- FIG6 is a schematic flow chart of a method for generating a personality-style database provided in an embodiment of the present application.
- FIG. 7 is a schematic diagram of using names as alignment labels of different style feature dimensions provided in an embodiment of the present application.
- FIG. 8 is another set of GUIs provided by an embodiment of the present application.
- FIG. 9 is another set of GUIs provided by an embodiment of the present application.
- FIG. 10 is another GUI provided by an embodiment of the present application.
- FIG. 11 is another schematic flowchart of the method for generating a virtual image provided in an embodiment of the present application.
- FIG. 12 is a schematic flowchart of an apparatus for generating a virtual image provided in an embodiment of the present application.
- prefixes such as “first” and “second” are used only to distinguish different description objects, and have no limiting effect on the position, order, priority, quantity or content of the described objects.
- the use of prefixes such as ordinal numbers to distinguish description objects in the embodiments of the present application does not constitute a limitation on the described objects.
- the meaning of "multiple" is two or more.
- virtual images are becoming more and more popular because they can provide users with diversified styles and intelligent interactive services.
- the terminal device can provide users with a variety of style parameters for users to customize. The more full the virtual image is, the more its style parameters increase dramatically. Since there are many style parameters and there is no correlation between the parameters, users need to have a certain knowledge base on the setting of virtual images in order to set up a virtual image that meets their expectations, which will increase the user's learning cost and affect the user's setting efficiency and user experience.
- the embodiment of the present application provides a method and device for generating a virtual image, which generates a virtual image that better meets the user's expectations by combining the user's historical information records, so that the user is more likely to resonate with the virtual image.
- the user does not need to choose from a large number of unrelated style parameters, which helps to reduce the user's learning cost when setting the virtual image, improves the efficiency of setting the virtual image, and also helps to improve the user's experience.
- FIG1 is a functional block diagram of a terminal device 100 provided in an embodiment of the present application.
- the terminal device 100 may include a text input unit 110, a display device 120, and a computing platform 130, wherein the text input unit 110 is used to obtain text content input by a user.
- the text input unit 110 may be used to obtain a nickname (or name) of a virtual image input by a user.
- the computing platform 130 may include one or more processors, such as processors 131 to 13n (n is a positive integer).
- the processor is a circuit with signal processing capabilities.
- the processor may be a circuit with instruction reading and execution capabilities, such as a central processing unit (CPU), a microprocessor, a graphics processing unit (GPU) (which can be understood as a microprocessor), or a digital signal processor (DSP); in another implementation, the processor can implement certain functions through the logical relationship of a hardware circuit, and the logical relationship of the hardware circuit is fixed or reconfigurable, such as a processor that is a hardware circuit implemented by an application-specific integrated circuit (ASIC) or a programmable logic device (PLD), such as a field programmable gate array (FPGA).
- ASIC application-specific integrated circuit
- PLD programmable logic device
- the process of the processor loading a configuration document to implement the hardware circuit configuration can be understood as the process of the processor loading instructions to implement the functions of some or all of the above units.
- the processor can also be a hardware circuit designed for artificial intelligence, which can be understood as an ASIC, such as a neural network processing unit (NPU), a tensor processing unit (TPU), a deep learning processing unit (DPU), etc.
- the computing platform 130 can also include a memory, the memory is used to store instructions, and some or all of the processors 131 to 13n can call the instructions in the memory and execute the instructions to achieve the corresponding functions.
- the computing platform 130 can obtain the text input unit
- the user receives the nickname (or name) of the virtual image sent by 110, and generates the virtual image according to the nickname (or name) of the virtual image and the historical information record of the user.
- the display device 120 in the vehicle cockpit is mainly divided into two categories.
- the first category is the vehicle display screen;
- the second category is a projection display screen, such as a head-up display (HUD).
- the vehicle display screen is a physical display screen and an important part of the vehicle infotainment system.
- There can be multiple display screens in the cockpit such as a digital instrument display screen, a central control screen, a display screen in front of the passenger in the co-pilot seat (also called the front passenger), a display screen in front of the left rear passenger, and a display screen in front of the right rear passenger.
- Even the window can be used as a display screen for display.
- Head-up display also known as a head-up display system.
- HUD includes, for example, a combiner-HUD (C-HUD) system, a windshield-HUD (W-HUD) system, and an augmented reality HUD (AR-HUD) system.
- C-HUD combiner-HUD
- W-HUD windshield-HUD
- AR-HUD augmented reality HUD
- the computing platform 130 when it generates a virtual image, it can display the virtual image to the user through the display device 120.
- the above display device 130 is described by taking a vehicle-mounted display screen and a projection display screen as examples, and the embodiments of the present application are not limited thereto.
- the display device 130 can also be a light display screen or a projection screen.
- Fig. 2 shows a schematic diagram of an exemplary display screen distribution in a vehicle cabin provided by an embodiment of the present application.
- the vehicle cabin may include a display screen 201 (or, it may also be referred to as a central control screen), a display screen 202 (or, it may also be referred to as a co-pilot entertainment screen), a display screen 203 (or, it may also be referred to as an entertainment screen in the left area of the second row), a display screen 204 (or, it may also be referred to as an entertainment screen in the right area of the second row) and an instrument screen.
- a display screen 201 or, it may also be referred to as a central control screen
- a display screen 202 or, it may also be referred to as a co-pilot entertainment screen
- a display screen 203 or, it may also be referred to as an entertainment screen in the left area of the second row
- a display screen 204 or, it may also be referred to as an entertainment screen in the right area of the second row
- FIG3 shows a schematic flow chart of a method 300 for generating a virtual image provided by an embodiment of the present application.
- the method 300 can be executed by the terminal device 100, or the method 300 can be executed by the computing platform 130, or the method 300 can be executed by a system on a chip (SoC) in the computing platform 130, or the method 300 can be executed by a processor in the computing platform 130.
- SoC system on a chip
- the following is an example of the method 300 being executed by a terminal device.
- the method 300 includes:
- the first text content may be a name, given name or nickname of the virtual image.
- the first text content may be the above-mentioned "Name A”.
- FIG4 shows a graphical user interface (GUI) provided in an embodiment of the present application.
- the GUI is a virtual image generation interface displayed by a display screen 201, and the virtual image generation interface includes a text input box 401, a virtual image generation control 402, and a cancel control 403.
- the terminal device can obtain the first text content "name A”.
- FIG4 is an example of a GUI displayed on a display screen 201 in a vehicle, but the present application is not limited thereto.
- the GUI may also be a display interface for generating a game character displayed on a mobile phone or computer, or may also be a display interface of an interactive robot.
- S320 Determine first personality characteristic information according to the first text content and the user's historical information record.
- the historical information record of the user includes the record of the user using the terminal device in the past period of time.
- the user's historical information records may include one or more of the following: records of the user searching for news in news applications, records of using efficiency applications (for example, applications for specifying work plans, voice-to-text applications, etc.), records of searching for movies or TV series in video applications, records of launching game applications, records of reading novels or jokes, records of searching in browsers, and records of listening to songs in music applications over the past period of time (for example, one month).
- efficiency applications for example, applications for specifying work plans, voice-to-text applications, etc.
- records of searching for movies or TV series in video applications records of launching game applications, records of reading novels or jokes, records of searching in browsers, and records of listening to songs in music applications over the past period of time (for example, one month).
- the user's historical information record may include the usage duration and frequency of different applications. For example, a user who frequently uses news applications may prefer a serious format text; a user who frequently uses video applications may prefer a novel format text; a user who frequently uses efficiency applications may prefer a concise and brief format text.
- the user's historical information record may be stored locally on the terminal device, or may be stored in other devices (e.g., a cloud server). If the user's historical information record is stored in the cloud server, when the first text content is obtained, the terminal device may request the user's historical information record from the cloud server. Thus, the terminal device may determine the first personality feature information based on the first text content and the historical information record.
- the first text content is "Name A”
- the user's historical information records include a record of launching a game application and using a game character named "Name A” in the game application.
- the terminal device can determine the personality characteristic information of the game character named "Name A" as the first personality characteristic information corresponding to the first text content.
- determining the first personality characteristic information based on the first text content and the user's historical information record includes: when the first text content is included in the historical information record, determining text co-occurrence words corresponding to the first text content based on the first text content; and determining the first personality characteristic information based on the text co-occurrence words.
- the terminal device can search the user's historical information record. For example, the terminal device can determine that the user has searched for "Mulan” twice through the browser in the past period of time and clicked on the search results related to the literary image in the search history. For example, the information in the search results is as follows:
- determining the first personality characteristic information according to the text co-occurring words includes: determining the first personality characteristic information corresponding to the first text content according to text similarities between the text co-occurring words and words related to different personality dimensions.
- the embedding layer of the bidirectional encoder representation from transformers (BERT) model can be used to extract word vectors for different words, or the word2Vec bag-of-words model can be used to extract word vectors.
- the following is an example of extracting word vectors using the embedding layer.
- V A (a 1 , a 2 ..., a i , ..., a N )
- V A represents the mathematical representation of word A in the word vector V A space
- V B represents the mathematical representation of word B in the word vector V B space.
- the similarity between word A and labeled word B can be calculated as shown in formula (1):
- Similarity AB is the similarity between word A and calibrated word B.
- n is the number of similarities calculated between the co-occurring words and the labeled words
- VnameA is the personality characteristic parameter of the name A
- VnameA can be used as the characteristic information of the name A in the personality dimension.
- Table 1 shows a Big Five personality dimension table.
- the calibrated words related to different personality dimensions may include “extrovert”, “talkative”, “quiet”, “shy”, “introvert”, “confident”, “dominant”, etc.
- determining the first personality characteristic information based on the first text content and the user's historical information record includes: when the first text content is not included in the historical information record, determining the similarity between the first text content and each of one or more names in the historical information record; and determining the first personality characteristic information based on the similarity of each name and the personality characteristic information corresponding to each name.
- the terminal device cannot extract the text co-occurrence related to the first text content. At this time, the terminal device can determine the first personality feature information in combination with the names of people similar to the first text content in the historical information record.
- the first personality characteristic information corresponding to the first text content can be determined by the following formula (3):
- V the first text content is the first personality characteristic information
- Similarity i is the similarity between the first text content and each person's name in the historical information record
- Vi is the personality characteristic parameter of each person's name
- N is the number of names in the historical information record.
- the first text content is "Hua Tielan", and "Hua Tielan” has not appeared in the historical information record, but includes names similar to "Hua Tielan", such as “Hua Mulan”, “Hua Rong”, “Hua Wuque” and “Temuzin”.
- the terminal device can calculate the similarity between "Hua Tielan” and the names that have appeared in the historical information record.
- the similarity of the names can be calculated by cosine similarity using the Embedding word vector, and the similarity is used as the weight.
- Table 2 shows the calculated similarity between "Hua Tielan” and the names that have appeared in the historical information record.
- VHuaTieLan is the personality characteristic parameter of "HuaTieLan” ( VHuaTieLan can be the personality characteristic information corresponding to "HuaTieLan”)
- VHuaMuLan is the personality characteristic parameter of "HuaMuLan”
- VHuaRong is the personality characteristic parameter of "HuaRong”
- VHuaWuQue is the personality characteristic parameter of "HuaWuQue” personality characteristic parameters
- V Temujin is the personality characteristic parameter of “Temujin”
- Similarity 1 is the similarity between “Hua Tielan” and “Hua Mulan” (for example, 0.4)
- Similarity 2 is the similarity between “Hua Tielan” and “Hua Rong” (for example, 0.2)
- Similarity 3 is the similarity between “Hua Tielan” and “Hua Wuque” (for example, 0.2)
- Similarity 4 is the similarity between “Hua Tielan” and “Te
- V Hua Mulan , V Hua Rong , V Hua Wu Que and V Temujin can refer to the calculation process of the personality characteristic parameter of name A in the above formula (2), which will not be repeated here.
- the above personality characteristic parameters may be a specific representation of the above personality characteristic information, and the embodiments of the present application are not limited thereto.
- the personality characteristic information may also be text content describing personality characteristics (e.g., "extrovert”, “talkative”, “introvert”, etc.).
- S330 Generate a virtual image corresponding to the first text content according to the first personality characteristic information.
- the terminal device may store a correspondence between personality characteristic information and a virtual image, and generate a virtual image corresponding to the first text content according to the first personality characteristic information, including: generating the virtual image according to the personality characteristic information and the correspondence.
- Table 3 shows a correspondence between personality feature information and a virtual image.
- the first text content is "Name A”.
- “Name A” Through “Name A” and the user's historical information record, it is determined that the personality characteristic information corresponding to Name A is extroversion and confidence. According to the corresponding relationship shown in Table 3 above, a virtual image 1 corresponding to "Name A" can be generated.
- determining the first personality characteristic information based on the first text content and the user's historical information records includes: prompting the user with multiple personality characteristic information based on the first text content and the historical information records; and determining the first personality characteristic information from the multiple personality characteristic information based on the user's input of the multiple personality characteristic information.
- FIG. 5 shows another GUI provided by an embodiment of the present application.
- the personality characteristic information corresponding to the "name A” can be determined based on the "name A” and the user's historical information record, wherein the historical information record indicates that the user has started Game 1 and used the character corresponding to the "name A", started the browser and browsed the search records related to the "name A", and searched for songs about the "name A” through the music application in the past week.
- the vehicle can display multiple personality characteristic information corresponding to the "name A” (for example, personality characteristic information 501-503) and the control corresponding to each personality characteristic information through the GUI based on the above historical information records.
- the personality characteristic information 501 corresponding to the "name A” in Game 1 is “extroverted, talkative and firm and confident”
- the personality characteristic information 502 corresponding to the "name A” in the browser's search record is "modest and magnanimous”
- the personality characteristic information 503 corresponding to the "name A” in the music application is "rude and suspicious”.
- the above user inputs the multiple personality characteristics information by clicking the control on the display screen as an example.
- the embodiments of the present application are not limited thereto.
- the user after seeing the personality characteristic information 501-503, the user can also issue a voice command "extrovert, talkative, and assertive".
- the vehicle After the vehicle obtains the voice command, it can determine that the personality characteristic information selected by the user is "extrovert, talkative, and assertive" according to the user's voice command, and thus can generate a corresponding virtual image according to the personality characteristic information.
- generating a virtual image corresponding to the first text content according to the first personality characteristic information includes: determining first style characteristic information according to the first personality characteristic information; and generating the virtual image according to the first style characteristic information.
- the terminal device stores a correspondence between personality feature information and style feature information, and determining the first style feature information according to the first personality feature information includes: determining the first style feature information according to the first personality feature information and the correspondence.
- the style feature information may be classified into different categories.
- Table 4 shows a classification method of the style feature information.
- Table 5 shows a correspondence between personality feature information and style feature information.
- the style characteristic information of the virtual image can be determined to be facial features 1, body shape 1, clothing 1, timbre 1, speaking speed 1, intonation 1, rhythm 1, response content 1, expression 1, gesture 1 and posture 1 based on the corresponding relationship shown in Table 5 above, so that a corresponding virtual image can be generated based on the style characteristic information.
- determining the first style characteristic information based on the first personality characteristic information includes: inputting the first personality characteristic information into a prediction model to obtain the first style characteristic information; wherein the prediction model is trained by sample training data, the sample training data includes a sample name, sample personality characteristic information and sample style characteristic information, and the sample style characteristic information includes at least one of an avatar, a body shape, a timbre, a speaking speed, a tone and an expression corresponding to the sample name.
- the above embodiment introduces the process of obtaining the personality characteristic parameters of a person's name through the text co-occurrence words of the name.
- the style features of other dimensions associated with the name can be used to label the personality characteristics of the name.
- the word vector of "Name A” and the personality characteristic parameters corresponding to the name are concatenated as input, and the shape and sound characteristics of "Name A" in "Movie 1" are used as labels to train the prediction model.
- data of multiple style characteristic dimensions with personality characteristic parameters as labels can be obtained. This type of data of multiple style characteristic dimensions can be used to obtain different style templates based on the style transfer method and store them in the prediction model.
- the above prediction model can also be understood as a personality-style database.
- FIG6 shows a schematic flow chart of a method 600 for generating a personality-style database provided in an embodiment of the present application.
- the method 600 may be executed by a device (e.g., a cloud server) including a model training device.
- the method 600 includes:
- the personality characteristic parameter of "name A” can be calculated according to the above formula (2).
- the personality characteristic parameters of "name A” appearing in different categories of historical information records may be different.
- Table 6 shows the correspondence between the text co-occurrence words and personality characteristic parameters of "name A" appearing in different categories of historical information records.
- FIG. 7 shows an example of an alignment method using names as alignment marks for different style feature dimensions provided by an embodiment of the present application.
- the avatar features of person A in “Movie 1” include avatar 1 and avatar 2
- the avatar features in “Game 1” include avatar 3 and avatar 4.
- the voice features of person A in “Movie 1” include voice feature 1, and the voice features in “Game 1” are voice feature 2.
- An association relationship can be established between person A, personality feature parameter 1, avatar 1, avatar 2, and voice feature 1, and an association relationship can be established between person A, personality feature parameter 2, avatar 3, avatar 4, and voice feature 2.
- generating a virtual image corresponding to the first text content according to the first style feature information includes: prompting the user with style feature information of one or more dimensions; and generating the virtual image according to the user's input of the style feature information of the one or more dimensions.
- style feature information of one or more dimensions is a style parameter list of one or more dimensions.
- FIG. 8 shows another set of GUIs provided by an embodiment of the present application.
- the vehicle can determine the personality feature information corresponding to the “name A” according to the “name A” and the user’s historical information record, wherein the user’s historical information record indicates that the user frequently opens game 1 and uses the character corresponding to the “name A” in the past week.
- the vehicle can determine the style parameter list of one or more dimensions corresponding to the “name A” according to the personality feature information corresponding to the “name A”, wherein the parameter list of the one or more dimensions includes an avatar list, a timbre list, and a tone list.
- the avatar list includes avatar 3 and avatar 4.
- the tone list includes voice 1 and voice 2, wherein the timbre of voice 1 is a timbre corresponding to the game character “name A”, and the timbre of voice 2 is another timbre corresponding to the game character “name A”.
- the tone list includes voice 3 and voice 4, wherein the tone of voice 3 is a tone corresponding to the game character “name A”, and the tone of voice 4 is another tone corresponding to the game character “name A”. Users can select their favorite style features from a list of style parameters in different dimensions.
- the vehicle when it is detected that the user has selected avatar 3, voice 1 and voice 3 and clicked the control 801, the vehicle can generate a virtual image 1 based on the avatar 3, the timbre corresponding to voice 1 and the pitch corresponding to voice 3, and display the virtual image 1 through the display screen.
- the avatar of the virtual image 1 is avatar 3 and the timbre and pitch of the voice signal "Hello, I am your virtual image Xiao A" emitted by the virtual image 1 match the timbre in the above-mentioned voice 1 and the pitch in voice 3 respectively.
- FIG8 is an example of allowing a user to manually select style parameters of different dimensions, but the embodiments of the present application are not limited thereto.
- the terminal device can automatically select the style parameter with the highest recommendation degree from the style parameter list of different dimensions as the style feature parameter of the virtual image, thereby automatically generating the corresponding virtual image.
- the terminal device can automatically randomly select a style parameter from the style parameter list of different dimensions as the style feature parameter of the virtual image, thereby automatically generating the corresponding virtual image. In this way, the process of allowing the user to select the style parameter of the virtual image is avoided, which helps to improve the user experience.
- the method 300 further includes: when it is detected that the user issues a voice command or a text input command, controlling the virtual image to make a voice reply or a text reply through a second text content, and the second text content is determined by the historical information record.
- FIG. 9 shows another set of GUIs provided by an embodiment of the present application.
- the user can perform voice interaction with the virtual image 1.
- the vehicle can detect the voice command “open the window” issued by the user.
- the vehicle can control the virtual image 1
- the user is responded to by voice through the lines in "Game 1".
- the lines of the game character "Name A" in "Game 1" include “Let me show you advanced operations”.
- the vehicle can control the virtual image to send a voice signal "Let me show you advanced operations” to respond to the user's voice command.
- using lines to replace the conventional response method during the conversation can give the user a more appropriate virtual image experience.
- generating a virtual image corresponding to the first text content according to the first personality characteristic information includes: prompting a plurality of virtual images to the user according to the first personality characteristic information; and determining the virtual image corresponding to the first text content from the plurality of virtual images according to the user's input to the plurality of virtual images.
- prompting a user with multiple virtual images based on the first personality characteristic information includes: determining first style characteristic information based on the first personality characteristic information; and prompting the user with multiple virtual images based on the first style characteristic information.
- FIG10 shows another GUI provided by an embodiment of the present application.
- the vehicle can control the central control screen to display the virtual image 1001 and the virtual image 1002 corresponding to "name A" according to the personality characteristic information corresponding to "name A”; or the vehicle can determine the style characteristic information corresponding to "name A” according to the personality characteristic information corresponding to "name A” and control the central control screen to display the virtual image 1001 and the virtual image 1002 corresponding to "name A” according to the style characteristic information.
- the central control screen can determine the style characteristic information corresponding to "name A” according to the personality characteristic information corresponding to "name A” and control the central control screen to display the virtual image 1001 and the virtual image 1002 corresponding to "name A” according to the style characteristic information.
- the user can also issue a voice command "select the virtual image on the far left". After the vehicle detects the voice command issued by the user, it can determine that the virtual image corresponding to the "name A" that the user wants to generate is virtual image 1001 according to the voice command.
- FIG11 shows a schematic flow chart of a method 1100 for generating a virtual image provided by an embodiment of the present application.
- the method 1100 may be executed by the terminal device 100, or the method 1100 may be executed by the computing platform 130, or the method 1100 may be executed by the SoC in the computing platform 130, or the method 1100 may be executed by the processor in the computing platform 130.
- the following description is made by taking the method 1100 executed by the terminal device as an example.
- the method 1100 includes:
- the terminal device when it is detected that the user inputs “name A” in the text input box 401 and clicks on the operation of generating a virtual image control 402 , the terminal device can obtain the information of “name A”.
- determining the personality characteristic parameters corresponding to the name based on the name and the user's historical information record includes: when the name is included in the user's historical information record, determining the personality characteristic parameters corresponding to the name based on text co-occurrence words of the name.
- the personality characteristic parameter can be calculated by the above formula (2).
- the personality characteristic parameter can be represented in the form of a vector.
- determining personality characteristic parameters corresponding to the name based on the name and the user's historical information record includes: when the name is not included in the user's historical information record, determining the similarity between the name and each of one or more names in the historical information record; and determining the personality characteristic parameters of the name based on the similarity of each name and the personality characteristic parameters corresponding to each name.
- the personality characteristic parameter may be input into the above prediction model (or personality-style database) to obtain the style characteristic parameter corresponding to the name.
- the terminal device can generate a virtual image based on the name input by the user and the user's historical information record. In this way, a virtual image that better meets the user's expectations can be generated based on the user's historical information record, and the user can resonate with the virtual image, which helps to improve the user's experience.
- Figure 12 shows a schematic block diagram of a device 1200 for generating a virtual image provided by an embodiment of the present application.
- the device 1200 includes: an acquisition unit 1210, used to acquire a first text content; a determination unit 1220, used to determine first personality feature information based on the first text content and the user's historical information record; and a virtual image generation unit 1230, used to generate a virtual image corresponding to the first text content based on the first personality feature information.
- the determination unit 1220 is used to: when the first text content is included in the historical information record, determine, based on the first text content, text co-occurrence words corresponding to the first text content; and determine the first personality characteristic information based on the text co-occurrence words.
- the determination unit 1220 is used to: when the first text content is not included in the historical information record, determine the similarity between the first text content and each of the one or more names in the historical information record; and determine the first personality characteristic information based on the similarity of each name and the personality characteristic information corresponding to each name.
- the device 1200 also includes a detection unit and a control unit, the detection unit is used to detect that the user issues a voice command or a text input command; the control unit is used to control the virtual image to make a voice reply or a text reply through a second text content, and the second text content is determined by the historical information record.
- the detection unit is used to detect that the user issues a voice command or a text input command
- the control unit is used to control the virtual image to make a voice reply or a text reply through a second text content, and the second text content is determined by the historical information record.
- the device 1200 also includes: a first prompting unit, used to prompt the user with multiple personality characteristic information based on the first text content and the historical information record; the determining unit 1220, used to determine the first personality characteristic information from the multiple personality characteristic information based on the user's input of the multiple personality characteristic information.
- a first prompting unit used to prompt the user with multiple personality characteristic information based on the first text content and the historical information record
- the determining unit 1220 used to determine the first personality characteristic information from the multiple personality characteristic information based on the user's input of the multiple personality characteristic information.
- the determination unit 1220 is further configured to determine first style characteristic information according to the first personality characteristic information; and the virtual image generation unit is configured to generate the virtual image according to the first style characteristic information.
- the determination unit 1220 is used to: input the first personality characteristic information into a prediction model to obtain the first style characteristic information; wherein the prediction model is trained by sample training data, the sample training data includes a sample name, sample personality characteristic information and sample style characteristic information, and the sample style characteristic information includes at least one of an avatar, shape, timbre, speaking speed, intonation and expression corresponding to the sample name.
- the device 1200 also includes: a second prompting unit, used to prompt the user with the style feature information of the one or more dimensions; and the virtual image generating unit 1230, used to generate the virtual image according to the user's input of the style feature information of the one or more dimensions.
- a second prompting unit used to prompt the user with the style feature information of the one or more dimensions
- the virtual image generating unit 1230 used to generate the virtual image according to the user's input of the style feature information of the one or more dimensions.
- the one or more dimensions of style feature information include at least one of avatar information, shape information, timbre information, speaking speed information, intonation information, and expression information.
- the device 1200 also includes: a third prompting unit, used to prompt a plurality of virtual images to the user according to the first personality characteristic information; and the virtual image generating unit 1230, used to determine a virtual image corresponding to the first text content from the plurality of virtual images according to the user's input to the plurality of virtual images.
- a third prompting unit used to prompt a plurality of virtual images to the user according to the first personality characteristic information
- the virtual image generating unit 1230 used to determine a virtual image corresponding to the first text content from the plurality of virtual images according to the user's input to the plurality of virtual images.
- the acquisition unit 1210 may be the computing platform in FIG. 1 or a processing circuit or a processing circuit in the computing platform. Taking the acquisition unit 1210 as the processor 131 in the computing platform as an example, the processor 131 can acquire the first text content input by the user.
- the determination unit 1220 is the computing platform in Figure 1 or a processing circuit, processor or controller in the computing platform. Taking the determination unit 1220 as the processor 132 in the computing platform as an example, the processor 132 can determine the first personality feature information based on the first text content obtained by the processor 131 and the historical information record of the user.
- the virtual image generation unit 1230 is the computing platform in Figure 1 or a processing circuit, processor or controller in the computing platform. Taking the virtual image generation unit 1230 as the processor 133 in the computing platform as an example, the processor 133 can generate a virtual image corresponding to the first text content according to the first personality feature information determined by the processor 132.
- the functions implemented by the above-mentioned acquisition unit 1210, the functions implemented by the determination unit 1220 and the functions implemented by the virtual image generation unit 1230 can be implemented by different processors, or, can be implemented by the same processor, or, part of the functions can be implemented by the same processor, and the embodiments of the present application are not limited to this.
- the division of the units in the above device is only a division of logical functions. In actual implementation, they can be fully or partially integrated into one physical entity, or they can be physically separated.
- the units in the device can be implemented in the form of a processor calling software; for example, the device includes a processor, the processor is connected to a memory, and instructions are stored in the memory.
- the processor calls the instructions stored in the memory to implement any of the above methods or realize the functions of the units of the device, wherein the processor is, for example, a general-purpose processor, such as a CPU or a microprocessor, and the memory is a memory in the device or a memory outside the device.
- the units in the device can be implemented in the form of hardware circuits, and the functions of some or all of the units can be realized by designing the hardware circuits.
- the hardware circuit can be understood as one or more processors; for example, in one implementation, the hardware circuit is an ASIC, and the functions of some or all of the above units are realized by designing the logical relationship of the components in the circuit; for example, in another implementation, the hardware circuit can be realized by PLD.
- FPGA as an example, it can include a large number of logic gate circuits, and the connection relationship between the logic gate circuits is configured through the configuration file, so as to realize the functions of some or all of the above units. All units of the above device may be implemented entirely in the form of a processor calling software, or entirely in the form of a hardware circuit, or partially in the form of a processor calling software and the rest in the form of a hardware circuit.
- Each unit in the above device may be one or more processors (or processing circuits) configured to implement the above method, such as a CPU, a GPU, an NPU, a TPU, a DPU, a microprocessor, a DSP, an ASIC, an FPGA, or a combination of at least two of these processor forms.
- processors or processing circuits configured to implement the above method, such as a CPU, a GPU, an NPU, a TPU, a DPU, a microprocessor, a DSP, an ASIC, an FPGA, or a combination of at least two of these processor forms.
- the SoC may include at least one processor for implementing any of the above methods or implementing the functions of each unit of the device.
- the type of the at least one processor may be different, for example, including CPU and FPGA, CPU and artificial intelligence processor, CPU and GPU, etc.
- An embodiment of the present application also provides a device, which includes a processing unit and a storage unit, wherein the storage unit is used to store instructions, and the processing unit executes the instructions stored in the storage unit so that the device executes the method or steps executed by the above embodiment.
- the processing unit may be the processor 131 - 13n shown in FIG. 1 .
- An embodiment of the present application further provides a system, which includes a computing platform and a display device, wherein the computing platform may include the above-mentioned device 1200.
- the display device may be a display screen, such as the display screens 201 - 204 described above.
- the embodiment of the present application further provides a terminal device, which may include the above-mentioned device 1200, or, Including the above system.
- the terminal device may be a mobile phone, a computer or a vehicle.
- the embodiment of the present application further provides a computer program product, which includes: a computer program code, and when the computer program code is executed on a computer, the computer executes the method in the above embodiment.
- the embodiment of the present application further provides a computer-readable medium, wherein the computer-readable medium stores a program code.
- the computer program code runs on a computer, the computer executes the method in the above embodiment.
- An embodiment of the present application further provides a chip, wherein the chip includes a circuit, and the circuit is used to execute the method in the above embodiment.
- each step of the above method can be completed by an integrated logic circuit of hardware in a processor or an instruction in the form of software.
- the method disclosed in conjunction with the embodiment of the present application can be directly embodied as a hardware processor for execution, or a combination of hardware and software modules in a processor for execution.
- the software module can be located in a storage medium mature in the art such as a random access memory, a flash memory, a read-only memory, a programmable read-only memory, or a power-on erasable programmable memory, a register, etc.
- the storage medium is located in a memory, and the processor reads the information in the memory and completes the steps of the above method in conjunction with its hardware. To avoid repetition, it is not described in detail here.
- the memory may include a read-only memory and a random access memory, and provide instructions and data to the processor.
- the size of the serial numbers of the above-mentioned processes does not mean the order of execution.
- the execution order of each process should be determined by its function and internal logic, and should not constitute any limitation on the implementation process of the embodiments of the present application.
- the disclosed systems, devices and methods can be implemented in other ways.
- the device embodiments described above are only schematic.
- the division of the units is only a logical function division. There may be other division methods in actual implementation, such as multiple units or components can be combined or integrated into another system, or some features can be ignored or not executed.
- Another point is that the mutual coupling or direct coupling or communication connection shown or discussed can be through some interfaces, indirect coupling or communication connection of devices or units, which can be electrical, mechanical or other forms.
- the units described as separate components may or may not be physically separated, and the components shown as units may or may not be physical units, that is, they may be located in one place or distributed on multiple network units. Some or all of the units may be selected according to actual needs to achieve the purpose of the solution of this embodiment.
- each functional unit in each embodiment of the present application may be integrated into one processing unit, or each unit may exist physically separately, or two or more units may be integrated into one unit.
- the functions are implemented in the form of software functional units and sold or used as independent products, they can be stored in a computer-readable storage medium.
- the part that makes technical contribution or the part of the technical solution can be embodied in the form of a software product, which is stored in a storage medium and includes several instructions for enabling a computer device (which can be a personal computer, server, or network device, etc.) to execute all or part of the steps of the method described in each embodiment of the present application.
- the aforementioned storage medium includes: U disk, mobile hard disk, read-only memory (ROM), random access memory (RAM), disk or optical disk, etc., various media that can store program codes.
Landscapes
- Physics & Mathematics (AREA)
- General Physics & Mathematics (AREA)
- Engineering & Computer Science (AREA)
- Theoretical Computer Science (AREA)
- User Interface Of Digital Computer (AREA)
Abstract
Description
V花铁兰=Similarity1*V花木兰+Similarity2*V花荣+Similarity3*V花无缺+Similarity4*V铁木真 (4)
Claims (24)
- 一种生成虚拟形象的方法,其特征在于,包括:获取第一文本内容;根据所述第一文本内容和用户的历史信息记录,确定第一人格特征信息;根据所述第一人格特征信息,生成所述第一文本内容对应的虚拟形象。
- 如权利要求1所述的方法,其特征在于,所述根据所述第一文本内容和用户的历史信息记录,确定所述第一人格特征信息,包括:在所述历史信息记录中包括所述第一文本内容时,根据所述第一文本内容,确定所述第一文本内容对应的文本共现词;根据所述文本共现词,确定所述第一人格特征信息。
- 如权利要求1所述的方法,其特征在于,所述根据所述第一文本内容和用户的历史信息记录,确定所述第一人格特征信息,包括:在所述历史信息记录中不包括所述第一文本内容时,确定所述第一文本内容与所述历史信息记录中一个或者多个人名中每个人名的相似度;根据所述每个人名的相似度以及所述每个人名对应的人格特征信息,确定所述第一人格特征信息。
- 如权利要求1至3中任一项所述的方法,其特征在于,所述方法还包括:在检测到用户发出语音指令或文本输入指令时,控制所述虚拟形象通过第二文本内容进行语音回复或者文本回复,所述第二文本内容由所述历史信息记录确定。
- 如权利要求1至4中任一项所述的方法,其特征在于,所述根据所述第一文本内容和用户的历史信息记录,确定第一人格特征信息,包括:根据所述第一文本内容和所述历史信息记录,向用户提示多个人格特征信息;根据用户对所述多个人格特征信息的输入,从所述多个人格特征信息中确定所述第一人格特征信息。
- 如权利要求1至5中任一项所述的方法,其特征在于,所述根据所述第一人格特征信息,生成所述第一文本内容对应的虚拟形象,包括:根据所述第一人格特征信息,确定第一风格特征信息;根据所述第一风格特征信息,生成所述虚拟形象。
- 如权利要求6所述的方法,其特征在于,所述根据所述第一人格特征信息,确定第一风格特征信息,包括:将所述第一人格特征信息输入预测模型中,得到所述第一风格特征信息;其中,所述预测模型由样本训练数据训练得到,所述样本训练数据包括样本人名、样本人格特征信息以及样本风格特征信息,所述样本风格特征信息包括与所述样本人名对应的头像、形体、音色、语速、语调、表情中的至少一项。
- 如权利要求6或7所述的方法,其特征在于,所述根据所述第一风格特征信息,生成所述虚拟形象,包括:向用户提示一个或者多个维度的风格特征信息;根据用户对所述一个或者多个维度的风格特征信息的输入,生成所述虚拟形象。
- 如权利要求8所述的方法,其特征在于,所述一个或者多个维度的风格特征信息包括头像信息、形体信息、音色信息、语速信息、语调信息、表情信息中的至少一项。
- 如权利要求1至9中任一项所述的方法,其特征在于,所述根据所述第一人格特征信息,生成所述第一文本内容对应的虚拟形象,包括:根据所述第一人格特征信息,向用户提示多个虚拟形象;根据用户对所述多个虚拟形象的输入,从所述多个虚拟形象中确定所述第一文本内容对应的虚拟形象。
- 一种生成虚拟形象的装置,其特征在于,包括:获取单元,用于获取第一文本内容;确定单元,用于根据所述第一文本内容和用户的历史信息记录,确定第一人格特征信息;虚拟形象生成单元,用于根据所述第一人格特征信息,生成所述第一文本内容对应的虚拟形象。
- 如权利要求11所述的装置,其特征在于,所述确定单元,用于:在所述历史信息记录中包括所述第一文本内容时,根据所述第一文本内容,确定所述第一文本内容对应的文本共现词;根据所述文本共现词,确定所述第一人格特征信息。
- 如权利要求11所述的装置,其特征在于,所述确定单元,用于:在所述历史信息记录中不包括所述第一文本内容时,确定所述第一文本内容与所述历史信息记录中一个或者多个人名中每个人名的相似度;根据所述每个人名的相似度以及所述每个人名对应的人格特征信息,确定所述第一人格特征信息。
- 如权利要求11至13中任一项所述的装置,其特征在于,所述装置还包括检测单元和控制单元,所述检测单元,用于检测到用户发出语音指令或文本输入指令;所述控制单元,用于控制所述虚拟形象通过第二文本内容进行语音回复或者文本回复,所述第二文本内容由所述历史信息记录确定。
- 如权利要求11至14中任一项所述的装置,其特征在于,所述装置还包括:第一提示单元,用于根据所述第一文本内容和所述历史信息记录,向用户提示多个人格特征信息;所述确定单元,用于根据用户对所述多个人格特征信息的输入,从所述多个人格特征信息中确定所述第一人格特征信息。
- 如权利要求11至15中任一项所述的装置,其特征在于,所述确定单元,还用于根据所述第一人格特征信息,确定第一风格特征信息;所述虚拟形象生成单元,用于根据所述第一风格特征信息,生成所述虚拟形象。
- 如权利要求16所述的装置,其特征在于,所述确定单元,用于:将所述第一人格特征信息输入预测模型中,得到所述第一风格特征信息;其中,所述预测模型由样本训练数据训练得到,所述样本训练数据包括样本人名、样 本人格特征信息以及样本风格特征信息,所述样本风格特征信息包括与所述样本人名对应的头像、形体、音色、语速、语调、表情中的至少一项。
- 如权利要求16或17所述的装置,其特征在于,所述装置还包括:第二提示单元,用于向用户提示一个或者多个维度的风格特征信息;所述虚拟形象生成单元,用于根据用户对所述一个或者多个维度的风格特征信息的输入,生成所述虚拟形象。
- 如权利要求18所述的装置,其特征在于,所述一个或者多个维度的风格特征信息包括头像信息、形体信息、音色信息、语速信息、语调信息、表情信息中的至少一项。
- 如权利要求11至19中任一项所述的装置,其特征在于,所述装置还包括:第三提示单元,用于根据所述第一人格特征信息,向用户提示多个虚拟形象;所述虚拟形象生成单元,用于根据用户对所述多个虚拟形象的输入,从所述多个虚拟形象中确定所述第一文本内容对应的虚拟形象。
- 一种生成虚拟形象的装置,其特征在于,包括:存储器,用于存储计算机程序;处理器,用于执行所述存储器中存储的计算机程序,以使得所述装置执行如权利要求1至10中任一项所述的方法。
- 一种终端设备,其特征在于,包括如权利要求11至21中任一项所述的装置。
- 一种计算机可读存储介质,其特征在于,其上存储有计算机程序,所述计算机程序被计算机执行时,以使得实现如权利要求1至10中任一项所述的方法。
- 一种芯片,其特征在于,所述芯片包括电路,所述电路用于执行如权利要求1至10中任一项所述的方法。
Priority Applications (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| CN202380089835.6A CN120457459A (zh) | 2023-02-28 | 2023-02-28 | 一种生成虚拟形象的方法和装置 |
| PCT/CN2023/078643 WO2024178590A1 (zh) | 2023-02-28 | 2023-02-28 | 一种生成虚拟形象的方法和装置 |
Applications Claiming Priority (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| PCT/CN2023/078643 WO2024178590A1 (zh) | 2023-02-28 | 2023-02-28 | 一种生成虚拟形象的方法和装置 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2024178590A1 true WO2024178590A1 (zh) | 2024-09-06 |
Family
ID=92589075
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/CN2023/078643 Ceased WO2024178590A1 (zh) | 2023-02-28 | 2023-02-28 | 一种生成虚拟形象的方法和装置 |
Country Status (2)
| Country | Link |
|---|---|
| CN (1) | CN120457459A (zh) |
| WO (1) | WO2024178590A1 (zh) |
Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN108510437A (zh) * | 2018-04-04 | 2018-09-07 | 科大讯飞股份有限公司 | 一种虚拟形象生成方法、装置、设备以及可读存储介质 |
| US10354256B1 (en) * | 2014-12-23 | 2019-07-16 | Amazon Technologies, Inc. | Avatar based customer service interface with human support agent |
| CN111339938A (zh) * | 2020-02-26 | 2020-06-26 | 广州腾讯科技有限公司 | 信息交互方法、装置、设备及存储介质 |
| US20200321020A1 (en) * | 2015-10-29 | 2020-10-08 | True Image Interactive, Inc. | Systems And Methods For Machine-Generated Avatars |
| CN113536007A (zh) * | 2021-07-05 | 2021-10-22 | 北京百度网讯科技有限公司 | 一种虚拟形象生成方法、装置、设备以及存储介质 |
Family Cites Families (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| KR102080878B1 (ko) * | 2019-02-07 | 2020-02-24 | 류경희 | 가상 현실 공간에서 서비스 제공을 위한 캐릭터 생성 및 학습 시스템 |
| CN110812843B (zh) * | 2019-10-30 | 2023-09-15 | 腾讯科技(深圳)有限公司 | 基于虚拟形象的交互方法及装置、计算机存储介质 |
| CN111309886B (zh) * | 2020-02-18 | 2023-03-21 | 腾讯科技(深圳)有限公司 | 一种信息交互方法、装置和计算机可读存储介质 |
| CN115317924B (zh) * | 2022-07-07 | 2025-06-10 | 网易(杭州)网络有限公司 | 游戏角色的生成方法、装置和电子设备 |
| CN115222857A (zh) * | 2022-07-27 | 2022-10-21 | 北京中电慧声科技有限公司 | 生成虚拟形象的方法、装置、电子设备和计算机可读介质 |
-
2023
- 2023-02-28 CN CN202380089835.6A patent/CN120457459A/zh active Pending
- 2023-02-28 WO PCT/CN2023/078643 patent/WO2024178590A1/zh not_active Ceased
Patent Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US10354256B1 (en) * | 2014-12-23 | 2019-07-16 | Amazon Technologies, Inc. | Avatar based customer service interface with human support agent |
| US20200321020A1 (en) * | 2015-10-29 | 2020-10-08 | True Image Interactive, Inc. | Systems And Methods For Machine-Generated Avatars |
| CN108510437A (zh) * | 2018-04-04 | 2018-09-07 | 科大讯飞股份有限公司 | 一种虚拟形象生成方法、装置、设备以及可读存储介质 |
| CN111339938A (zh) * | 2020-02-26 | 2020-06-26 | 广州腾讯科技有限公司 | 信息交互方法、装置、设备及存储介质 |
| CN113536007A (zh) * | 2021-07-05 | 2021-10-22 | 北京百度网讯科技有限公司 | 一种虚拟形象生成方法、装置、设备以及存储介质 |
Also Published As
| Publication number | Publication date |
|---|---|
| CN120457459A (zh) | 2025-08-08 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US12462441B2 (en) | Iterative image generation from text | |
| US20250342828A1 (en) | Digital assistant control of applications | |
| US20190163768A1 (en) | Automatically curated image searching | |
| CN113330455A (zh) | 使用有条件的生成对抗网络查找互补的数字图像 | |
| WO2024022437A1 (zh) | 一种显示方法、装置和移动载体 | |
| CN115408611A (zh) | 菜单推荐方法、装置、计算机设备和存储介质 | |
| CN113486260A (zh) | 互动信息的生成方法、装置、计算机设备及存储介质 | |
| CN115858850A (zh) | 内容推荐方法、装置、车辆及计算机可读存储介质 | |
| CN118276746A (zh) | 用于图像编辑的方法、装置、设备、介质和程序产品 | |
| JP7289756B2 (ja) | 生成装置、生成方法および生成プログラム | |
| WO2024178590A1 (zh) | 一种生成虚拟形象的方法和装置 | |
| CN120804187A (zh) | 视觉数据生成方法、装置、电子设备及可读存储介质 | |
| CN116166823A (zh) | 基于用户偏好的多媒体信息展示方法及装置、存储介质 | |
| WO2025036359A1 (zh) | 图像处理方法、装置、计算机设备及存储介质 | |
| WO2025039700A1 (zh) | 一种搜索结果排序方法、装置及系统 | |
| CN118409685A (zh) | 车机界面的显示方法、装置、电子设备、车辆及存储介质 | |
| CN116797322A (zh) | 提供商品对象信息的方法及电子设备 | |
| CN112734949B (zh) | Vr内容的属性修改方法、装置、计算机设备及存储介质 | |
| CN114741602A (zh) | 对象推荐方法、目标模型的训练方法、装置及设备 | |
| CN114816038A (zh) | 虚拟现实内容生成方法、装置及计算机可读存储介质 | |
| US20240119489A1 (en) | Product score unique to user | |
| WO2025102351A1 (zh) | 语音交互方法和装置 | |
| KR20260029177A (ko) | 보조 미디어 컨텐츠를 제공하는 방법 및 이를 수행하는 전자 장치 | |
| CN121329550A (zh) | 商品推荐方法及系统 | |
| CN109167723B (zh) | 图像的处理方法、装置、存储介质及电子设备 |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 23924565 Country of ref document: EP Kind code of ref document: A1 |
|
| WWE | Wipo information: entry into national phase |
Ref document number: 202380089835.6 Country of ref document: CN |
|
| WWP | Wipo information: published in national office |
Ref document number: 202380089835.6 Country of ref document: CN |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 23924565 Country of ref document: EP Kind code of ref document: A1 |