WO2023051651A1 - 音乐生成方法、装置、设备、存储介质及程序 - Google Patents
音乐生成方法、装置、设备、存储介质及程序 Download PDFInfo
- Publication number
- WO2023051651A1 WO2023051651A1 PCT/CN2022/122334 CN2022122334W WO2023051651A1 WO 2023051651 A1 WO2023051651 A1 WO 2023051651A1 CN 2022122334 W CN2022122334 W CN 2022122334W WO 2023051651 A1 WO2023051651 A1 WO 2023051651A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- scale
- target
- music
- action
- information
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Images
Classifications
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10H—ELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
- G10H1/00—Details of electrophonic musical instruments
- G10H1/0008—Associated control or indicating means
- G10H1/0025—Automatic or semi-automatic music composition, e.g. producing random music, applying rules from music theory or modifying a musical piece
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10H—ELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
- G10H1/00—Details of electrophonic musical instruments
- G10H1/0008—Associated control or indicating means
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10H—ELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
- G10H2210/00—Aspects or methods of musical processing having intrinsic musical character, i.e. involving musical theory or musical parameters or relying on musical knowledge, as applied in electrophonic musical tools or instruments
- G10H2210/031—Musical analysis, i.e. isolation, extraction or identification of musical elements or musical parameters from a raw acoustic signal or from an encoded audio signal
- G10H2210/081—Musical analysis, i.e. isolation, extraction or identification of musical elements or musical parameters from a raw acoustic signal or from an encoded audio signal for automatic key or tonality recognition, e.g. using musical rules or a knowledge base
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10H—ELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
- G10H2210/00—Aspects or methods of musical processing having intrinsic musical character, i.e. involving musical theory or musical parameters or relying on musical knowledge, as applied in electrophonic musical tools or instruments
- G10H2210/101—Music Composition or musical creation; Tools or processes therefor
- G10H2210/111—Automatic composing, i.e. using predefined musical rules
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10H—ELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
- G10H2210/00—Aspects or methods of musical processing having intrinsic musical character, i.e. involving musical theory or musical parameters or relying on musical knowledge, as applied in electrophonic musical tools or instruments
- G10H2210/101—Music Composition or musical creation; Tools or processes therefor
- G10H2210/131—Morphing, i.e. transformation of a musical piece into a new different one, e.g. remix
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10H—ELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
- G10H2210/00—Aspects or methods of musical processing having intrinsic musical character, i.e. involving musical theory or musical parameters or relying on musical knowledge, as applied in electrophonic musical tools or instruments
- G10H2210/395—Special musical scales, i.e. other than the 12-interval equally tempered scale; Special input devices therefor
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10H—ELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
- G10H2210/00—Aspects or methods of musical processing having intrinsic musical character, i.e. involving musical theory or musical parameters or relying on musical knowledge, as applied in electrophonic musical tools or instruments
- G10H2210/555—Tonality processing, involving the key in which a musical piece or melody is played
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10H—ELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
- G10H2220/00—Input/output interfacing specifically adapted for electrophonic musical tools or instruments
- G10H2220/155—User input interfaces for electrophonic musical instruments
- G10H2220/201—User input interfaces for electrophonic musical instruments for movement interpretation, i.e. capturing and recognizing a gesture or a specific kind of movement, e.g. to control a musical instrument
-
- G—PHYSICS
- G10—MUSICAL INSTRUMENTS; ACOUSTICS
- G10H—ELECTROPHONIC MUSICAL INSTRUMENTS; INSTRUMENTS IN WHICH THE TONES ARE GENERATED BY ELECTROMECHANICAL MEANS OR ELECTRONIC GENERATORS, OR IN WHICH THE TONES ARE SYNTHESISED FROM A DATA STORE
- G10H2220/00—Input/output interfacing specifically adapted for electrophonic musical tools or instruments
- G10H2220/155—User input interfaces for electrophonic musical instruments
- G10H2220/441—Image sensing, i.e. capturing images or optical patterns for musical purposes or musical control purposes
- G10H2220/455—Camera input, e.g. analyzing pictures from a video camera and using the analysis results as control data
Definitions
- Embodiments of the present disclosure relate to the technical field of artificial intelligence, and in particular to a music generation method, device, device, electronic device, computer-readable storage medium, computer program product, and computer program.
- piano keys may be displayed on a screen of a terminal device, and a user may simulate playing a piano by touching the keys on the screen, thereby realizing music creation.
- Embodiments of the present disclosure provide a music generation method, device, device, electronic device, computer-readable storage medium, computer program product, and computer program, so as to solve the problem of poor interactivity in music creation methods.
- an embodiment of the present disclosure provides a method for generating music, including:
- action information of a first action performed by the first user within the preset duration according to the first video where the action information includes an action type and an action duration
- target music corresponding to the preset duration is generated.
- an embodiment of the present disclosure provides a music generating device, including:
- An acquisition module configured to acquire the first video obtained by collecting the first user within a preset duration
- a determining module configured to determine action information of a first action performed by the first user within the preset duration according to the first video, the action information including action type and action duration;
- a generating module configured to generate target music corresponding to the preset duration according to the action information of the first action.
- an embodiment of the present disclosure provides an electronic device, including: a processor and a memory;
- the memory stores computer-executable instructions
- the processor executes the computer-executed instructions to implement the music generation method in the first aspect and various possible implementation manners of the first aspect.
- an embodiment of the present disclosure provides a computer-readable storage medium, where computer-executable instructions are stored in the computer-readable storage medium, and when the processor executes the computer-executable instructions, the above first aspect and the first A music generation method in various possible implementations of the aspect.
- an embodiment of the present disclosure provides a computer program product, including a computer program.
- the computer program is executed by a processor, the music generation method in the above first aspect and various possible implementation manners of the first aspect is implemented.
- an embodiment of the present disclosure provides a computer program, when the computer program is executed by a processor, implements the music generation method in the above first aspect and various possible implementation manners of the first aspect.
- the music generation method, device, device, electronic device, computer-readable storage medium, computer program product and computer program provided by the embodiments of the present disclosure includes: obtaining the first user's collected data within a preset time period A video, according to the first video, determine the action information of the first action performed by the first user within the preset duration, the action information includes the action type and action duration, and then according to the first action action information to generate the target music corresponding to the preset duration.
- FIG. 1 is a schematic diagram of an application scenario provided by an embodiment of the present disclosure
- FIG. 2 is a schematic flow diagram of a music generation method provided by an embodiment of the present disclosure
- FIG. 3 is a schematic flowchart of another music generation method provided by an embodiment of the present disclosure.
- FIG. 4 is a schematic diagram of a display interface provided by an embodiment of the present disclosure.
- FIG. 5 is a schematic diagram of another display interface provided by an embodiment of the present disclosure.
- FIG. 6 is a schematic diagram of another display interface provided by an embodiment of the present disclosure.
- FIG. 7 is a schematic structural diagram of a music generating device provided by an embodiment of the present disclosure.
- FIG. 8 is a schematic structural diagram of an electronic device provided by an embodiment of the present disclosure.
- Scale Arranging the tones in the mode in a ladder form from low to high (called ascending) or from high to low (called descending) from the beginning of the tonic to the end of the tonic is called a scale.
- the natural heptatonic scale is the most widely used heptatonic scale. Its interval organization is that there are 5 whole tones in each octave, which are divided into two strings and three strings, and the two strings are separated by semitones. .
- the 7 scales are: 1(Do), 2(Re), 3(Mi), 4(Fa), 5(So), 6(La), 7(Si).
- the embodiment of the present disclosure does not limit the scale system, and may be applied to any scale system.
- the natural seven-tone scale is used as an example when referring to illustrations.
- Timbre Different sounds always have distinctive characteristics in terms of waveforms, and different objects vibrate with different characteristics. Different sounding bodies have different timbres due to their different materials and structures. For example, the sounds produced by pianos and violins are different from those produced by people, and the sounds produced by each person are also different. Therefore, timbre can be understood as the characteristic of sound.
- FIG. 1 is a schematic diagram of an application scenario provided by an embodiment of the present disclosure.
- this application scenario involves terminal equipment.
- the terminal equipment is provided with a video acquisition device such as a camera.
- the terminal equipment is also provided with an audio playback device such as a loudspeaker.
- the terminal device is also provided with a music generating device.
- the music generation device may be in the form of software and/or hardware.
- the music generation device may be a processor, a chip, a chip module, a module, a unit, an application program, etc. in a terminal device.
- the image capture device can capture user actions (such as body movements, facial expressions, gestures, etc.) to obtain videos.
- the image acquisition device sends the video to the music generation device.
- the music generating means can generate music based on the video.
- the music generation device may recognize a user's action sequence from a video, and map the user's action into a musical scale to obtain a musical scale sequence. In turn, music can be generated from sequences of scales.
- the music generating device sends the generated music to the audio playing device, and the audio playing device plays the music.
- the user can make the terminal device generate different music to realize music creation.
- the user can perform a series of actions to interact with the terminal device to realize music creation, which makes the interaction and interest higher, and can improve the user experience.
- the terminal device in the embodiments of the present disclosure may be any electronic device with a video acquisition device and an audio playback device, including but not limited to: smart phones, tablet computers, notebook computers, smart TVs, smart wearable devices, smart home devices, etc.
- Fig. 2 is a schematic flowchart of a method for generating music provided by an embodiment of the present disclosure.
- the method in this embodiment may be executed by a terminal device, or may be executed by a music generating apparatus in the terminal device.
- the method of this embodiment includes:
- S201 Acquire a first video obtained by collecting a first user within a preset time period.
- This embodiment is applied to a scenario where the first user creates target music.
- the preset duration represents the execution granularity of this embodiment in the time dimension.
- the first video captured within a preset duration is acquired, and the target music corresponding to the first preset duration is generated based on the first video.
- the preset duration may also be referred to as a time window.
- the preset duration can be set to a small value, for example, 1 second, 100 milliseconds, and so on.
- the preset duration may be the duration for the user to perform an action.
- the user usually needs to repeat this embodiment for many times.
- the corresponding preset durations may be the same or different. In this way, the target music corresponding to multiple preset durations is combined into the final complete music.
- Method 1 The music generation process is carried out synchronously with the video capture process. That is to say, the method of this embodiment is executed synchronously during the process of capturing the video for the first user, so as to generate music based on the video while capturing the video. Specifically, with reference to Fig. 1, whenever the video capture device captures a first video with a preset duration, it sends the first video to the music generation device, and the music generation device generates the target music corresponding to the preset duration based on the first video. .
- Mode 2 the music generation process and the video capture process are executed asynchronously.
- the video capture device captures the user to obtain a second video, for example, the duration of the second video may be 3 minutes or 5 minutes.
- the second video may be a complete video of the user dancing a dance.
- the second video collected above is stored in a preset storage space.
- the music generation device obtains the second video from the preset storage space, and takes the preset duration as the time granularity, and obtains the first video corresponding to the preset duration from the second video, according to the first The video generates target music corresponding to the preset duration.
- S202 According to the first video, determine action information of a first action performed by the first user within the preset duration, where the action information includes an action type and an action duration.
- the first action is an action performed by a preset body part of the first user.
- the first action may be a body action, such as raising a leg, raising an arm, twisting, bending, and the like.
- the first action may be a gesture action, for example: an applause gesture, an OK gesture, a scissors gesture, and the like.
- the first action may also be an emoticon, such as: smiling emoticon, laughing emoticon, pouting emoticon, surprised emoticon, and the like.
- the action information of the first action may be determined in the following manner: determine a target body part, where the target body part includes at least one of the following parts of the first user: limbs, hands, face ; respectively detecting feature information of the target body part in multiple image frames of the first video; determining the first action according to the feature information of the target body part detected in the multiple image frames action information.
- the action information of the first action includes: the action type of the first action and the action duration of the first action.
- the action duration of the first action may refer to the hold time of the first action, or refer to the sum of the time taken to complete the first action and the hold time of the first action.
- the target body part is a limb as an example.
- the first video includes 12 image frames
- the target recognition algorithm uses the target recognition algorithm to recognize the limbs in the image frame to obtain the feature information of the limbs.
- the left arm hangs naturally in image frame 1
- the angle between the left arm and the torso in image frame 2 is 30 degrees
- the angle between the left arm and the torso in image frame 3 is 60 degrees
- the left arm in image frame 4 is 60 degrees.
- the angle between the arm and the torso is 90 degrees
- the angle between the left arm and the torso in image frame 5 is 120 degrees.
- the included angle between the left arm and the torso in image frame 6 to image frame 12 is 180 degrees.
- the action type of the first action is determined as "raise the left arm” according to the feature information of the limbs recognized from the above image frames.
- the time interval between image frame 1 and image frame 12 may be used as the action duration of the first action, or the time interval between image frame 6 and image frame 12 may be taken as the action duration of the first action.
- the first video can be sampled according to a preset sampling rule, and the above-mentioned identification process of the target body part can be performed on each image frame after sampling, which can improve the performance of the second action.
- the real-time performance of the detection result of an action can be improved.
- S203 Generate target music corresponding to the preset duration according to the action information of the first action.
- a piece of music is usually formed by a variety of musical elements, such as: scales, timbres, tunes, etc.
- the correspondence between different actions and music elements may be defined in advance. In this way, using the above correspondence, the first action can be mapped to the relevant information of the music element, so as to generate the target music.
- a correspondence between body movements and musical scales may be defined, that is, different types of body movements correspond to different types of musical scales.
- a correspondence between facial expressions and timbres may be defined, that is, different facial expressions correspond to different timbres.
- a correspondence between gesture actions and tunes may be defined, that is, different gesture actions correspond to different tunes. It should be noted that the above correspondences are only some possible examples. The different examples above can be used in combination with each other.
- the following will take the mapping of different actions to different scales as an example for illustration, and the scale information of the target scale corresponding to the first action may be determined according to the action information of the first action.
- Information includes scale type and duration.
- the scale type of the target scale may be determined according to the action type of the first action and a preset corresponding relationship, and the preset corresponding relationship is used to indicate the corresponding relationship between different action types and different scale types .
- the preset corresponding relationship may be as shown in Table 1.
- the sound length of the target scale may be determined according to the action duration of the first action.
- the target music may be generated according to the scale information of the target scale.
- the method may further include: playing the target music.
- the first user can hear the effect of the target music in time.
- the first user can adjust the action in real time to correct the target music when he thinks that the effect of the target music is not good by listening to the target music Or adjust to improve the efficiency of users creating music and improve user experience.
- the music generation method provided in this embodiment includes: acquiring a first video obtained by collecting the first user within a preset time period, and determining the first user’s time in the preset time period according to the first video.
- the action information of the first action executed within the first action the action information includes the action type and the action duration, and according to the action information of the first action, the target music corresponding to the preset duration is generated.
- the user can perform a series of actions to interact with the terminal device to realize music creation, which makes the interactivity and interest high, and can improve the user experience.
- Fig. 3 is a schematic flowchart of another method for generating music provided by an embodiment of the present disclosure. As shown in Figure 3, the method of this embodiment includes:
- the target music is the music to be composed/generated.
- the timbre type of the target music can be any one of the following: piano timbre, accordion timbre, violin timbre, harmonica timbre, and so on.
- the generation mode is the first generation mode based on free creation, or the second generation mode based on reference music.
- the first user can freely organize the sequence of actions and the action duration of each action. That is to say, the first user is not restricted during the music creation process, and can create music completely according to his own preference.
- the first user needs to organize the sequence of actions based on the scale sequence in the reference music, and adjust the duration of each scale in the reference music through the action duration of each action, so as to generate the target music.
- the generated target music is equivalent to adapting the rhythm of the reference music.
- the terminal device may receive the timbre type of the target music input by the first user, and receive the generation mode of the target music input by the first user.
- FIG. 4 is a schematic diagram of a display interface provided by an embodiment of the present disclosure.
- the terminal device can display the first interface as shown in Figure 4(a) to the user, in the first interface provide options for multiple timbre types, and the user can select the target in the first interface according to the creative needs
- the tone type of the music For example, suppose the user selects "piano tone" in the first interface. After the user clicks the next step, the terminal device may display the second interface as shown in FIG. 4( b ).
- the terminal device provides the option of generating mode in the second interface, and the user can select the generating mode of the target music in the second interface according to the creation requirement.
- S303 Obtain a first video obtained by collecting the first user within a preset time period.
- S304 According to the first video, determine action information of a first action performed by the first user within the preset duration, where the action information includes an action type and an action duration.
- S305 Determine, according to the action information of the first action, scale information of a target scale corresponding to the first action, where the scale information includes a scale type and a sound length.
- S303 to S307 may be repeatedly executed for multiple rounds.
- the user performs the first action, and the terminal device determines the scale information of the target scale according to the action information of the first action.
- the musical scale information of the target musical scale determined during multiple rounds of execution forms the target music created by the user.
- FIG. 5 is a schematic diagram of another display interface provided by an embodiment of the present disclosure.
- the terminal device displays the third interface as shown in FIG. 5( b ).
- the current creation progress is displayed on the third interface.
- the target scale determined by the terminal device is "1"; assuming that the second action performed by the user is “raise the right leg”, the terminal device The determined target scale is “2"; assuming that the third action performed by the user is “raise the right arm (over the shoulder)", the target scale determined by the terminal device is "4"; assuming that the user performs the fourth action The action is “raise the left arm (over the shoulder)", and the target scale determined by the terminal device is "5"; as shown in Figure 5(b), the current creation progress is "1 2 4 5".
- the creation progress can be displayed in various ways, for example, in the form of musical scale sequence, numbered musical notation, etc., which is not limited in this embodiment.
- the terminal device receives the completion instruction and determines that the creation of the target music is completed.
- the reference music is the music that needs to be adapted to generate the target music.
- the reference music may be specified by the user, or may be randomly determined by the terminal device.
- S310 Perform scale analysis processing on the reference music to obtain a scale sequence corresponding to the reference music.
- the scale sequence includes multiple reference scales, and the multiple reference scales are arranged in sequence according to the order in which they appear in the reference music.
- the terminal device may receive reference music input by the first user.
- FIG. 6 is a schematic diagram of another display interface provided by an embodiment of the present disclosure.
- the terminal device displays the fourth interface as shown in FIG. 6( b ).
- the terminal device displays selection controls for the user to select reference music on the fourth interface.
- the user can input reference music to the terminal device by selecting the control.
- the terminal device can also display an input control on the fourth interface, so that the user can also input the name of the reference music to the terminal device through the input control.
- the user can also input reference music to the terminal device by voice. This embodiment does not limit it.
- the terminal device performs scale analysis processing on the reference music to obtain
- the reference scales form the following scale sequence in the order in which the reference scales appear.
- the terminal device may display the fifth interface shown in FIG. 6(c), in which the above-mentioned scale sequence is displayed. In this way, the user can perform actions corresponding to the reference scales according to the order of the reference scales in the scale sequence displayed on the fifth interface.
- S311 Determine a target reference scale according to the order of the reference scales in the scale sequence.
- S311 to S318 may be repeatedly executed for multiple rounds.
- the first reference scale in the scale sequence is determined as the target reference scale.
- the second reference scale in the scale sequence is determined as the target reference scale. and so on.
- the user needs to perform an action corresponding to the target reference scale.
- the target reference scale can also be highlighted (for example, the target reference scale is located in the rectangular box in Figure 6(c)), so that it is more intuitive Remind the user of the current creation progress and the actions that need to be performed.
- S312 Acquire a first video obtained by collecting the first user within a preset time period.
- S313 According to the first video, determine action information of a first action performed by the first user within a preset duration, where the action information includes an action type and an action duration.
- S314 According to the action information of the first action, determine the scale information of the target scale corresponding to the first action, where the scale information includes a scale type and a sound length.
- S315 Determine whether the scale type of the target scale is the same as the scale type of the target reference scale.
- the user when the user selects the second generating mode, the user needs to perform corresponding actions according to the order of the reference scales in the reference music. Therefore, in each round of execution, it is necessary to determine whether the scale type of the target scale is the same as the scale type of the target reference scale. If they are the same, S316 can be executed. If not, return to S312 to re-detect the actions performed by the user.
- the terminal device may also display a prompt message on the fifth interface shown in FIG. 6(c), such as "the current action is incorrect, Please perform the action again" to prompt the user to make timely adjustments.
- S317 Play the target music corresponding to the preset duration.
- S318 Determine whether the target reference scale is the last scale in the scale sequence.
- the scale sequence in the generated target music is the same as the scale sequence in the reference music.
- the difference between the two is that the length of each scale is different.
- the rhythm of the music is different. Therefore, the target music can be regarded as obtained by adapting the rhythm of the reference music.
- the user can perform a series of actions to interact with the terminal device to realize music creation, so that interactivity and interest are high, and user experience can be improved. Furthermore, the user can also use the first generation mode based on free creation, or the second generation mode based on reference music to create music, which further improves the fun of creating music.
- Fig. 7 is a schematic structural diagram of a music generating device provided by an embodiment of the present disclosure.
- the means may be in the form of software and/or hardware.
- the apparatus may be a terminal device, or a processor, chip, chip module, module, unit, application program, etc. integrated into the terminal device.
- the music generation device 700 provided in this embodiment includes: an acquisition module 701 , a determination module 702 and a generation module 703 .
- the obtaining module 701 is configured to obtain the first video obtained by collecting the first user within a preset duration
- a determining module 702 configured to determine action information of a first action performed by the first user within the preset duration according to the first video, where the action information includes an action type and an action duration;
- a generating module 703, configured to generate target music corresponding to the preset duration according to the action information of the first action.
- the generating module 703 is specifically used to:
- the scale information of the target scale corresponding to the first action determines the scale information of the target scale corresponding to the first action, and the scale information includes a scale type and a sound length;
- the target music is generated according to the scale information of the target scale.
- the generating module 703 is specifically used to:
- the pitch length of the target scale is determined according to the action duration of the first action.
- the obtaining module 701 is also used to obtain the generation mode of the target music, the generation mode is the first generation mode based on free creation, or the second generation mode based on reference music;
- the generation module 703 is specifically configured to: generate the target music according to the generation mode of the target music and the scale information of the target scale.
- the generating module 703 is specifically used to:
- the generation mode of the target music is the first generation mode, then generate the target music according to the scale information of the target scale; or,
- the generation mode of the target music is the second generation mode, then determine the target reference scale from the reference music, and when the scale type of the target scale is the same as the scale type of the target reference scale, according to the The scale information of the target scale is used to generate the target music.
- the acquiring module 701 is further configured to: acquire the reference music, perform scale analysis processing on the reference music, and obtain a scale sequence corresponding to the reference music, and the scale sequence includes multiple reference music Scales, the plurality of reference scales are arranged sequentially according to the order in which they appear in the reference music;
- the generating module 703 is specifically configured to: determine the target reference scale according to the order of the reference scales in the scale sequence.
- the generating module 703 is specifically used to:
- the target music is generated according to the tone color type and scale information of the target scale.
- the device further includes:
- the playing module is used to play the target music.
- the determining module 702 is specifically configured to:
- the target body part comprising at least one of the following parts of the first user: limbs, hands, face;
- Action information of the first action is determined according to the feature information of the target body part detected in the plurality of image frames.
- the music generating device provided in this embodiment can be used to execute the music generating method provided in any of the above method embodiments, and its implementation principle and technical effect are similar, and will not be repeated here.
- the embodiments of the present disclosure further provide an electronic device.
- the electronic device 800 may be a terminal device or a server.
- the terminal equipment may include but not limited to mobile phones, notebook computers, digital broadcast receivers, personal digital assistants (Personal Digital Assistant, PDA for short), tablet computers (Portable Android Device, PAD for short), portable multimedia players (Portable Mobile terminals such as Media Player (PMP for short), vehicle-mounted terminals (such as vehicle navigation terminals), etc., and fixed terminals such as digital television (Digital Television, digital TV for short), desktop computers, etc.
- PDA Personal Digital Assistant
- PDA Personal Digital Assistant
- PAD Personal Android Device
- portable multimedia players Portable Mobile terminals such as Media Player (PMP for short
- vehicle-mounted terminals such as vehicle navigation terminals
- fixed terminals such as digital television (Digital Television, digital TV for short), desktop computers, etc.
- the electronic device shown in FIG. 8 is only an example, and should not limit the functions and scope of use of the embodiments of the present disclosure.
- an electronic device 800 may include a processing device (such as a central processing unit, a graphics processing unit, etc.) 808 loads programs in random access memory (Random Access Memory, RAM for short) 803 to execute various appropriate actions and processes. In the RAM 803, various programs and data necessary for the operation of the electronic device 800 are also stored.
- the processing device 801, ROM 802, and RAM 803 are connected to each other through a bus 804.
- An input/output (Input/Output, I/O for short) interface 805 is also connected to the bus 804 .
- an input device 806 including, for example, a touch screen, a touchpad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc.; ), a speaker, a vibrator, etc.
- a storage device 808 including, for example, a magnetic tape, a hard disk, etc.
- the communication means 809 may allow the electronic device 800 to communicate with other devices wirelessly or by wire to exchange data. While FIG. 8 shows electronic device 800 having various means, it is to be understood that implementing or having all of the means shown is not a requirement. More or fewer means may alternatively be implemented or provided.
- embodiments of the present disclosure include a computer program product, which includes a computer program carried on a computer-readable medium, where the computer program includes program codes for executing the methods shown in the flowcharts.
- the computer program may be downloaded and installed from a network via communication means 809, or from storage means 808, or from ROM 802.
- the processing device 801 the above-mentioned functions defined in the methods of the embodiments of the present disclosure are executed.
- the above-mentioned computer-readable medium in the present disclosure may be a computer-readable signal medium or a computer-readable storage medium or any combination of the above two.
- a computer readable storage medium may be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof.
- Computer-readable storage media may include, but are not limited to, electrical connections with one or more wires, portable computer diskettes, hard disks, random access memory (RAM), read-only memory (ROM), erasable Programming read-only memory (Erasable Programmable Read Only Memory, referred to as EPROM or flash memory), optical fiber, portable compact disk read-only memory (Compact Disk Read Only Memory, referred to as CD-ROM), optical storage device, magnetic storage device, or any of the above the right combination.
- a computer-readable storage medium may be any tangible medium that contains or stores a program that can be used by or in conjunction with an instruction execution system, apparatus, or device.
- a computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave carrying computer-readable program code therein. Such propagated data signals may take many forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the foregoing.
- a computer-readable signal medium may also be any computer-readable medium other than a computer-readable storage medium, which can transmit, propagate, or transmit a program for use by or in conjunction with an instruction execution system, apparatus, or device .
- the program code contained on the computer readable medium can be transmitted by any appropriate medium, including but not limited to: electric wire, optical cable, radio frequency (Radio Frequency, RF for short), etc., or any suitable combination of the above.
- the above-mentioned computer-readable medium may be included in the above-mentioned electronic device, or may exist independently without being incorporated into the electronic device.
- the above-mentioned computer-readable medium carries one or more programs, and when the above-mentioned one or more programs are executed by the electronic device, the electronic device is made to execute the methods shown in the above-mentioned embodiments.
- Computer program code for carrying out the operations of the present disclosure can be written in one or more programming languages, or combinations thereof, including object-oriented programming languages—such as Java, Smalltalk, C++, and conventional Procedural Programming Language - such as "C" or a similar programming language.
- the program code may execute entirely on the user's computer, partly on the user's computer, as a stand-alone software package, partly on the user's computer and partly on a remote computer or entirely on the remote computer or server.
- the remote computer can be connected to the user's computer through any kind of network, including a Local Area Network (LAN) or a Wide Area Network (WAN), or it can be connected to an external A computer (connected via the Internet, eg, using an Internet service provider).
- LAN Local Area Network
- WAN Wide Area Network
- each block in a flowchart or block diagram may represent a module, program segment, or portion of code that contains one or more logical functions for implementing specified executable instructions.
- the functions noted in the block may occur out of the order noted in the figures. For example, two blocks shown in succession may, in fact, be executed substantially concurrently, or they may sometimes be executed in the reverse order, depending upon the functionality involved.
- each block of the block diagrams and/or flowchart illustrations, and combinations of blocks in the block diagrams and/or flowchart illustrations can be implemented by a dedicated hardware-based system that performs the specified functions or operations , or may be implemented by a combination of dedicated hardware and computer instructions.
- the units involved in the embodiments described in the present disclosure may be implemented by software or by hardware. Wherein, the name of the unit does not constitute a limitation of the unit itself under certain circumstances, for example, the first obtaining unit may also be described as "a unit for obtaining at least two Internet Protocol addresses".
- exemplary types of hardware logic components include: Field Programmable Gate Array (Field Programmable Gate Array, FPGA for short), Application Specific Integrated Circuit (ASIC for short), Application Specific Standard Products ( Application Specific Standard Parts (ASSP for short), System on Chip (SOC for short), Complex Programmable Logic Device (CPLD for short), etc.
- a machine-readable medium may be a tangible medium that may contain or store a program for use by or in conjunction with an instruction execution system, apparatus, or device.
- a machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium.
- a machine-readable medium may include, but is not limited to, electronic, magnetic, optical, electromagnetic, infrared, or semiconductor systems, apparatus, or devices, or any suitable combination of the foregoing.
- machine-readable storage media would include one or more wire-based electrical connections, portable computer discs, hard drives, random access memory (RAM), read only memory (ROM), erasable programmable read only memory (EPROM or flash memory), optical fiber, compact disk read only memory (CD-ROM), optical storage, magnetic storage, or any suitable combination of the foregoing.
- RAM random access memory
- ROM read only memory
- EPROM or flash memory erasable programmable read only memory
- CD-ROM compact disk read only memory
- magnetic storage or any suitable combination of the foregoing.
- a music generation method including:
- action information of a first action performed by the first user within the preset duration according to the first video where the action information includes an action type and an action duration
- target music corresponding to the preset duration is generated.
- generating the target music corresponding to the preset duration includes:
- the scale information of the target scale corresponding to the first action determines the scale information of the target scale corresponding to the first action, and the scale information includes a scale type and a sound length;
- the target music is generated according to the scale information of the target scale.
- determining the scale information of the target scale corresponding to the first action includes:
- the pitch length of the target scale is determined according to the action duration of the first action.
- before obtaining the first video obtained by capturing the first user within a preset time period further includes:
- the generation mode is a first generation mode based on free creation, or a second generation mode based on reference music;
- generating the target music according to the scale information of the target scale includes:
- the target music is generated according to the generation mode of the target music and the scale information of the target scale.
- generating the target music according to the generation mode of the target music and the scale information of the target scale includes:
- the generation mode of the target music is the first generation mode, then generate the target music according to the scale information of the target scale; or,
- the generation mode of the target music is the second generation mode, then determine the target reference scale from the reference music, and when the scale type of the target scale is the same as the scale type of the target reference scale, according to the The scale information of the target scale is used to generate the target music.
- the target reference scale from the reference music before determining the target reference scale from the reference music, it also includes:
- determining the target reference scale from the reference music includes:
- the target reference scale is determined according to the order of the reference scales in the scale sequence.
- generating the target music according to the scale information of the target scale includes:
- the target music is generated according to the tone color type and scale information of the target scale.
- after generating the target music according to the scale information of the target scale further includes:
- determining the action information of the first action performed by the first user within the preset duration includes:
- the target body part comprising at least one of the following parts of the first user: limbs, hands, face;
- Action information of the first action is determined according to the feature information of the target body part detected in the plurality of image frames.
- a music generation device including:
- An acquisition module configured to acquire the first video obtained by collecting the first user within a preset duration
- a determining module configured to determine action information of a first action performed by the first user within the preset duration according to the first video, the action information including action type and action duration;
- a generating module configured to generate target music corresponding to the preset duration according to the action information of the first action.
- the generating module is specifically configured to:
- the scale information of the target scale corresponding to the first action determines the scale information of the target scale corresponding to the first action, and the scale information includes a scale type and a sound length;
- the target music is generated according to the scale information of the target scale.
- the generating module is specifically configured to:
- the pitch length of the target scale is determined according to the action duration of the first action.
- the acquisition module is further configured to acquire the generation mode of the target music, the generation mode is the first generation mode based on free creation, or the first generation mode based on reference music Two generation mode;
- the generation module is specifically configured to: generate the target music according to the generation mode of the target music and the scale information of the target scale.
- the generating module is specifically configured to:
- the generation mode of the target music is the first generation mode, then generate the target music according to the scale information of the target scale; or,
- the generation mode of the target music is the second generation mode, then determine the target reference scale from the reference music, and when the scale type of the target scale is the same as the scale type of the target reference scale, according to the The scale information of the target scale is used to generate the target music.
- the acquiring module is further configured to acquire the reference music, perform scale analysis processing on the reference music, and obtain a scale sequence corresponding to the reference music, in the scale sequence including a plurality of reference scales, the plurality of reference scales are arranged sequentially according to the order of their appearance in the reference music;
- the generating module is specifically configured to determine the target reference scale according to the order of the reference scales in the scale sequence.
- the generating module is specifically configured to:
- the target music is generated according to the tone color type and scale information of the target scale.
- the device further includes:
- the playing module is used to play the target music.
- the determining module is specifically configured to:
- the target body part comprising at least one of the following parts of the first user: limbs, hands, face;
- Action information of the first action is determined according to the feature information of the target body part detected in the plurality of image frames.
- an electronic device including: at least one processor and a memory;
- the memory stores computer-executable instructions
- the at least one processor executes the computer-executed instructions stored in the memory, so that the at least one processor executes the music generation method described in the above first aspect and various possible implementation manners of the first aspect.
- a computer-readable storage medium stores computer-executable instructions, and when a processor executes the computer-executable instructions, Realize the music generation method described in the above first aspect and various possible implementation manners of the first aspect.
- a computer program product including a computer program, when the computer program is executed by a processor, various possible implementations of the first aspect and the first aspect can be realized The music generation method described in the manner.
- a computer program is provided, and when the computer program is executed by a processor, the music described in the first aspect and various possible implementation modes of the first aspect is realized. generate method.
Landscapes
- Engineering & Computer Science (AREA)
- Physics & Mathematics (AREA)
- Acoustics & Sound (AREA)
- Multimedia (AREA)
- Theoretical Computer Science (AREA)
- User Interface Of Digital Computer (AREA)
Abstract
Description
| 动作类型 | 音阶类型 |
| 抬起左腿 | 1(Do) |
| 抬起右腿 | 2(Re) |
| 抬起左胳膊 | 3(Mi) |
| 抬起右胳膊 | 4(Fa) |
| 抬起左胳膊过肩 | 5(So) |
| 抬起右胳膊过肩 | 6(La) |
| T字型(即左右胳膊与肩平齐,左右腿竖直站立) | 7(Si) |
Claims (14)
- 一种音乐生成方法,包括:获取在预设时长内对第一用户进行采集得到的第一视频;根据所述第一视频,确定所述第一用户在所述预设时长内执行的第一动作的动作信息,所述动作信息包括动作类型和动作时长;根据所述第一动作的动作信息,生成所述预设时长对应的目标音乐。
- 根据权利要求1所述的方法,其中,根据所述第一动作的动作信息,生成所述预设时长对应的目标音乐,包括:根据所述第一动作的动作信息,确定所述第一动作对应的目标音阶的音阶信息,所述音阶信息包括音阶类型和音长;根据所述目标音阶的音阶信息,生成所述目标音乐。
- 根据权利要求2所述的方法,其中,根据所述第一动作的动作信息,确定所述第一动作对应的目标音阶的音阶信息,包括:根据所述第一动作的动作类型和预设对应关系,确定所述目标音阶的音阶类型,所述预设对应关系用于指示不同动作类型与不同音阶类型之间的对应关系;根据所述第一动作的动作时长,确定所述目标音阶的音长。
- 根据权利要求2或3所述的方法,其中,获取在预设时长内对第一用户进行采集得到的第一视频之前,还包括:获取所述目标音乐的生成模式,所述生成模式为基于自由创作的第一生成模式,或者,为基于参考音乐的第二生成模式;相应的,根据所述目标音阶的音阶信息,生成所述目标音乐,包括:根据所述目标音乐的生成模式,以及所述目标音阶的音阶信息,生成所述目标音乐。
- 根据权利要求4所述的方法,其中,根据所述目标音乐的生成模式,以及所述目标音阶的音阶信息,生成所述目标音乐,包括:若所述目标音乐的生成模式为所述第一生成模式,则根据所述目标音阶的音阶信息,生成所述目标音乐;或者,若所述目标音乐的生成模式为所述第二生成模式,则从所述参考音乐中确定目标参考音阶,在所述目标音阶的音阶类型与所述目标参考音阶的音阶类型相同时,根据所述目标音阶的音阶信息,生成所述目标音乐。
- 根据权利要求5所述的方法,其中,从所述参考音乐中确定目标参考音阶之前,还包括:获取所述参考音乐;对所述参考音乐进行音阶解析处理,得到所述参考音乐对应的音阶序列,所述音阶序列中包括多个参考音阶,所述多个参考音阶按照各自在所述参考音乐中出现顺序依次排列;相应的,从所述参考音乐中确定目标参考音阶,包括:按照所述音阶序列中各参考音阶的顺序,确定所述目标参考音阶。
- 根据权利要求2至6中任一项所述的方法,其中,根据所述目标音阶的音阶信息,生成所述目标音乐,包括:获取所述目标音乐的音色类型;根据所述音色类型和所述目标音阶的音阶信息,生成所述目标音乐。
- 根据权利要求2至7中任一项所述的方法,其中,根据所述目标音阶的音阶信息,生成所述目标音乐之后,还包括:播放所述目标音乐。
- 根据权利要求1至8中任一项所述的方法,其中,根据所述第一视频,确定所述第一用户在所述预设时长内执行的第一动作的动作信息,包括:确定目标身体部位,所述目标身体部位包括所述第一用户的下述部位中的至少一种:肢体、手部、面部;在所述第一视频的多个图像帧中分别检测所述目标身体部位的特征信息;根据所述多个图像帧中检测得到的所述目标身体部位的特征信息,确定所述第一动作的动作信息。
- 一种音乐生成装置,包括:获取模块,用于获取在预设时长内对第一用户进行采集得到的第一视频;确定模块,用于根据所述第一视频,确定所述第一用户在所述预设时长内执行的第一动作的动作信息,所述动作信息包括动作类型和动作时长;生成模块,用于根据所述第一动作的动作信息,生成所述预设时长对应的目标音乐。
- 一种电子设备,包括:处理器和存储器;所述存储器存储计算机执行指令;所述处理器执行所述计算机执行指令,实现如权利要求1至9中任一项所述的音乐生成方法。
- 一种计算机可读存储介质,所述计算机可读存储介质中存储有计算机执行指令,当处理器执行所述计算机执行指令时,实现如权利要求1至9中任一项所述的音乐生成方法。
- 一种计算机程序产品,包括计算机程序,所述计算机程序被处理器执行时实现如权利要求1至9中任一项所述的音乐生成方法。
- 一种计算机程序,所述计算机程序被处理器执行时实现如权利要求1至9中任一项所述的音乐生成方法。
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US18/571,086 US20240290305A1 (en) | 2021-09-28 | 2022-09-28 | Music generation method and apparatus, device, storage medium, and program |
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| CN202111140485.0A CN115881064A (zh) | 2021-09-28 | 2021-09-28 | 音乐生成方法、装置、设备、存储介质及程序 |
| CN202111140485.0 | 2021-09-28 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2023051651A1 true WO2023051651A1 (zh) | 2023-04-06 |
Family
ID=85763262
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/CN2022/122334 Ceased WO2023051651A1 (zh) | 2021-09-28 | 2022-09-28 | 音乐生成方法、装置、设备、存储介质及程序 |
Country Status (3)
| Country | Link |
|---|---|
| US (1) | US20240290305A1 (zh) |
| CN (1) | CN115881064A (zh) |
| WO (1) | WO2023051651A1 (zh) |
Cited By (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN118283343A (zh) * | 2024-03-27 | 2024-07-02 | 北京度友信息技术有限公司 | 基于配乐的视频生成方法、装置以及设备 |
Families Citing this family (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| EP4660994A4 (en) * | 2024-04-24 | 2025-12-10 | Beijing Zitiao Network Technology Co Ltd | Music generation method, music generation apparatus, and computer readable storage medium |
| CN119668456A (zh) * | 2024-12-13 | 2025-03-21 | 北京字跳网络技术有限公司 | 一种媒体内容生成方法、装置、设备、介质及程序产品 |
Citations (8)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN103885663A (zh) * | 2014-03-14 | 2014-06-25 | 深圳市东方拓宇科技有限公司 | 一种生成和播放音乐的方法及其对应终端 |
| CN105700808A (zh) * | 2016-02-18 | 2016-06-22 | 广东欧珀移动通信有限公司 | 音乐播放方法、装置及终端设备 |
| CN109119057A (zh) * | 2018-08-30 | 2019-01-01 | Oppo广东移动通信有限公司 | 音乐创作方法、装置及存储介质和穿戴式设备 |
| CN109413351A (zh) * | 2018-10-26 | 2019-03-01 | 平安科技(深圳)有限公司 | 一种音乐生成方法及装置 |
| CN110827789A (zh) * | 2019-10-12 | 2020-02-21 | 平安科技(深圳)有限公司 | 音乐生成方法、电子装置及计算机可读存储介质 |
| CN110874171A (zh) * | 2018-08-31 | 2020-03-10 | 阿里巴巴集团控股有限公司 | 音频信息处理方法及装置 |
| CN110944085A (zh) * | 2019-11-12 | 2020-03-31 | 南京邮电大学 | 一种晃动智能手机产生音乐的方法 |
| CN112698757A (zh) * | 2020-12-25 | 2021-04-23 | 北京小米移动软件有限公司 | 界面交互方法、装置、终端设备及存储介质 |
Family Cites Families (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN108053831A (zh) * | 2017-12-05 | 2018-05-18 | 广州酷狗计算机科技有限公司 | 音乐生成、播放、识别方法、装置及存储介质 |
| WO2020154422A2 (en) * | 2019-01-22 | 2020-07-30 | Amper Music, Inc. | Methods of and systems for automated music composition and generation |
| CN111276122B (zh) * | 2020-01-14 | 2023-10-27 | 广州酷狗计算机科技有限公司 | 音频生成方法及装置、存储介质 |
| WO2021159203A1 (en) * | 2020-02-10 | 2021-08-19 | 1227997 B.C. Ltd. | Artificial intelligence system & methodology to automatically perform and generate music & lyrics |
| CN112927665B (zh) * | 2021-01-22 | 2022-08-30 | 咪咕音乐有限公司 | 创作方法、电子设备和计算机可读存储介质 |
-
2021
- 2021-09-28 CN CN202111140485.0A patent/CN115881064A/zh active Pending
-
2022
- 2022-09-28 WO PCT/CN2022/122334 patent/WO2023051651A1/zh not_active Ceased
- 2022-09-28 US US18/571,086 patent/US20240290305A1/en active Pending
Patent Citations (8)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN103885663A (zh) * | 2014-03-14 | 2014-06-25 | 深圳市东方拓宇科技有限公司 | 一种生成和播放音乐的方法及其对应终端 |
| CN105700808A (zh) * | 2016-02-18 | 2016-06-22 | 广东欧珀移动通信有限公司 | 音乐播放方法、装置及终端设备 |
| CN109119057A (zh) * | 2018-08-30 | 2019-01-01 | Oppo广东移动通信有限公司 | 音乐创作方法、装置及存储介质和穿戴式设备 |
| CN110874171A (zh) * | 2018-08-31 | 2020-03-10 | 阿里巴巴集团控股有限公司 | 音频信息处理方法及装置 |
| CN109413351A (zh) * | 2018-10-26 | 2019-03-01 | 平安科技(深圳)有限公司 | 一种音乐生成方法及装置 |
| CN110827789A (zh) * | 2019-10-12 | 2020-02-21 | 平安科技(深圳)有限公司 | 音乐生成方法、电子装置及计算机可读存储介质 |
| CN110944085A (zh) * | 2019-11-12 | 2020-03-31 | 南京邮电大学 | 一种晃动智能手机产生音乐的方法 |
| CN112698757A (zh) * | 2020-12-25 | 2021-04-23 | 北京小米移动软件有限公司 | 界面交互方法、装置、终端设备及存储介质 |
Cited By (2)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN118283343A (zh) * | 2024-03-27 | 2024-07-02 | 北京度友信息技术有限公司 | 基于配乐的视频生成方法、装置以及设备 |
| CN118283343B (zh) * | 2024-03-27 | 2025-01-21 | 北京度友信息技术有限公司 | 基于配乐的视频生成方法、装置以及设备 |
Also Published As
| Publication number | Publication date |
|---|---|
| US20240290305A1 (en) | 2024-08-29 |
| CN115881064A (zh) | 2023-03-31 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| US11158102B2 (en) | Method and apparatus for processing information | |
| US20210029305A1 (en) | Method and apparatus for adding a video special effect, terminal device and storage medium | |
| US11749246B2 (en) | Systems and methods for music simulation via motion sensing | |
| US11514923B2 (en) | Method and device for processing music file, terminal and storage medium | |
| WO2023051651A1 (zh) | 音乐生成方法、装置、设备、存储介质及程序 | |
| CN111798821B (zh) | 声音转换方法、装置、可读存储介质及电子设备 | |
| WO2020119150A1 (zh) | 节奏点识别方法、装置、电子设备及存储介质 | |
| CN112380362B (zh) | 基于用户交互的音乐播放方法、装置、设备及存储介质 | |
| JP2022505118A (ja) | 画像処理方法、装置、ハードウェア装置 | |
| WO2021129628A1 (zh) | 视频特效处理方法及装置 | |
| EP4604064A1 (en) | Image processing method and apparatus, electronic device, and storage medium | |
| CN115691544A (zh) | 虚拟形象口型驱动模型的训练及其驱动方法、装置和设备 | |
| CN111833460A (zh) | 增强现实的图像处理方法、装置、电子设备及存储介质 | |
| WO2020151491A1 (zh) | 图像形变的控制方法、装置和硬件装置 | |
| WO2023061229A1 (zh) | 视频生成方法及设备 | |
| CN115774539B (zh) | 和声处理方法、装置、设备及介质 | |
| WO2023160713A1 (zh) | 音乐生成方法、装置、设备、存储介质及程序 | |
| KR102637788B1 (ko) | 이동 단말기용 악기 장치 및 그의 제어 방법 및 프로그램 | |
| EP4597488A1 (en) | Audio processing method and apparatus, and electronic device | |
| CN116737994A (zh) | 视频、唱谱音频和曲谱同步播放方法、装置、设备和介质 | |
| WO2025194881A1 (zh) | 音频播放方法、装置及终端设备 | |
| CN118567468A (zh) | 交互方法、装置、设备及存储介质 | |
| CN107404581B (zh) | 移动终端的乐器模拟方法、装置及存储介质和移动终端 | |
| CN116932812A (zh) | 乐谱更新方法、装置、电子设备和计算机可读介质 | |
| CN115138062A (zh) | 设备游戏互动方法、电子设备及可读存储介质 |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 22875034 Country of ref document: EP Kind code of ref document: A1 |
|
| WWE | Wipo information: entry into national phase |
Ref document number: 18571086 Country of ref document: US |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 32PN | Ep: public notification in the ep bulletin as address of the adressee cannot be established |
Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205A DATED 05.07.2024) |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 22875034 Country of ref document: EP Kind code of ref document: A1 |