WO2020259130A1 - 精选片段处理方法、装置、电子设备及可读介质 - Google Patents

精选片段处理方法、装置、电子设备及可读介质 Download PDF

Info

Publication number
WO2020259130A1
WO2020259130A1 PCT/CN2020/091071 CN2020091071W WO2020259130A1 WO 2020259130 A1 WO2020259130 A1 WO 2020259130A1 CN 2020091071 W CN2020091071 W CN 2020091071W WO 2020259130 A1 WO2020259130 A1 WO 2020259130A1
Authority
WO
WIPO (PCT)
Prior art keywords
segment
video
audio
preset duration
selected segment
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/CN2020/091071
Other languages
English (en)
French (fr)
Inventor
何尧
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Beijing ByteDance Network Technology Co Ltd
Original Assignee
Beijing ByteDance Network Technology Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Beijing ByteDance Network Technology Co Ltd filed Critical Beijing ByteDance Network Technology Co Ltd
Publication of WO2020259130A1 publication Critical patent/WO2020259130A1/zh
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Images

Classifications

    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/40Client devices specifically adapted for the reception of or interaction with content, e.g. set-top-box [STB]; Operations thereof
    • H04N21/43Processing of content or additional data, e.g. demultiplexing additional data from a digital video stream; Elementary client operations, e.g. monitoring of home network or synchronising decoder's clock; Client middleware
    • H04N21/431Generation of visual interfaces for content selection or interaction; Content or additional data rendering
    • H04N21/4312Generation of visual interfaces for content selection or interaction; Content or additional data rendering involving specific graphical features, e.g. screen layout, special fonts or colors, blinking icons, highlights or animations
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/40Client devices specifically adapted for the reception of or interaction with content, e.g. set-top-box [STB]; Operations thereof
    • H04N21/43Processing of content or additional data, e.g. demultiplexing additional data from a digital video stream; Elementary client operations, e.g. monitoring of home network or synchronising decoder's clock; Client middleware
    • H04N21/433Content storage operation, e.g. storage operation in response to a pause request, caching operations
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/40Client devices specifically adapted for the reception of or interaction with content, e.g. set-top-box [STB]; Operations thereof
    • H04N21/43Processing of content or additional data, e.g. demultiplexing additional data from a digital video stream; Elementary client operations, e.g. monitoring of home network or synchronising decoder's clock; Client middleware
    • H04N21/439Processing of audio elementary streams
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/40Client devices specifically adapted for the reception of or interaction with content, e.g. set-top-box [STB]; Operations thereof
    • H04N21/43Processing of content or additional data, e.g. demultiplexing additional data from a digital video stream; Elementary client operations, e.g. monitoring of home network or synchronising decoder's clock; Client middleware
    • H04N21/44Processing of video elementary streams, e.g. splicing a video clip retrieved from local storage with an incoming video stream or rendering scenes according to encoded video stream scene graphs
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/40Client devices specifically adapted for the reception of or interaction with content, e.g. set-top-box [STB]; Operations thereof
    • H04N21/47End-user applications
    • H04N21/472End-user interface for requesting content, additional data or services; End-user interface for interacting with content, e.g. for content reservation or setting reminders, for requesting event notification, for manipulating displayed content
    • H04N21/47205End-user interface for requesting content, additional data or services; End-user interface for interacting with content, e.g. for content reservation or setting reminders, for requesting event notification, for manipulating displayed content for manipulating displayed content, e.g. interacting with MPEG-4 objects, editing locally
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/40Client devices specifically adapted for the reception of or interaction with content, e.g. set-top-box [STB]; Operations thereof
    • H04N21/47End-user applications
    • H04N21/482End-user interface for programme selection
    • H04N21/4825End-user interface for programme selection using a list of items to be played back in a given order, e.g. playlists
    • HELECTRICITY
    • H04ELECTRIC COMMUNICATION TECHNIQUE
    • H04NPICTORIAL COMMUNICATION, e.g. TELEVISION
    • H04N21/00Selective content distribution, e.g. interactive television or video on demand [VOD]
    • H04N21/80Generation or processing of content or additional data by content creator independently of the distribution process; Content per se
    • H04N21/85Assembly of content; Generation of multimedia applications
    • H04N21/854Content authoring
    • H04N21/8549Creating video summaries, e.g. movie trailer

Definitions

  • the embodiments of the present disclosure relate to the field of Internet technology, for example, to a method, device, electronic device, and readable medium for processing selected fragments.
  • K song applications users can watch audio and video released by other users, or select their favorite songs to record audio and video and publish them.
  • the user can enter the name of the favorite song in the search box, and then click any K song option in the search results to enter the singing interface of the song to record the song. And usually, after the song is recorded, the user can choose to record the video image, and then can generate and save the audio and video based on the audio and video images of the recorded song, and directly save the saved after the user’s release instruction is detected
  • the audio and video of is posted on the video list interface for other users to watch.
  • the embodiments of the present disclosure provide a selected fragment processing method, device, electronic device, and readable medium, which help users quickly capture key points, thereby improving user experience.
  • the embodiment of the present disclosure provides a selected fragment processing method, which includes:
  • the recorded audio content and image content of the target song generate the audio and video of the target song
  • the selected segment is added to the video list interface for display.
  • the embodiment of the present disclosure also provides a selected fragment processing device, which includes:
  • the audio and video generation module is set to generate the audio and video of the target song based on the recorded audio content and image content of the target song;
  • the selected segment generating module is configured to save the audio and video when a saving event is detected, and determine a segment with a preset duration from the audio and video as the selected segment;
  • a display module is added, which is set to add the selected fragments to the video list interface for display when a publishing event is detected.
  • An embodiment of the present disclosure also provides an electronic device, which includes:
  • One or more processors are One or more processors;
  • Memory set to store one or more programs
  • the one or more processors When the one or more programs are executed by the one or more processors, the one or more processors implement the selected segment processing method according to any embodiment of the present disclosure.
  • the embodiment of the present disclosure provides a readable medium on which a computer program is stored.
  • the computer program is executed by a processor, the selected fragment processing method as described in any embodiment of the present disclosure is implemented.
  • FIG. 1 shows a flowchart of a method for processing selected fragments according to an embodiment of the present disclosure
  • Fig. 2 shows a flowchart of another selected fragment processing method provided by an embodiment of the present disclosure
  • Fig. 3 shows a flowchart of another selected fragment processing method provided by an embodiment of the present disclosure
  • Fig. 4 shows a flowchart of another selected fragment processing method provided by an embodiment of the present disclosure
  • FIG. 5 shows a schematic structural diagram of a selected fragment processing apparatus provided by an embodiment of the present disclosure
  • Fig. 6 shows a schematic structural diagram of an electronic device provided by an embodiment of the present disclosure.
  • each embodiment provides optional features and examples at the same time. Multiple features recorded in the embodiments can be combined to form multiple alternative solutions. Each numbered embodiment should not be combined Only regarded as a technical solution.
  • first and second mentioned in the present disclosure are only used to distinguish different devices, modules or units, and are not used to limit the order or interdependence of the functions performed by these devices, modules or units relationship.
  • Figure 1 shows a flow chart of a method for processing selected fragments provided by an embodiment of the present disclosure.
  • the display of audio and video in the video list interface is not convenient for users to quickly capture the key points, which in turn leads to poor user experience.
  • the method may be executed by the selected segment processing apparatus or electronic device provided in the embodiments of the present disclosure, and the apparatus may be implemented by software and/or hardware.
  • the so-called electronic device may be a server device that carries the function of processing selected fragments, or a terminal device that configures the K song application provided by the server.
  • the selected fragment processing method provided in the embodiment of the present disclosure includes S110 to S130.
  • S110 Generate audio and video of the target song according to the recorded audio content and image content of the target song.
  • the target song refers to the audio and/or image song selected by the user to be recorded.
  • the interface In the interface, record the image content; 2) In the audio and video synchronization recording interface, record the audio content and image content of the target song at the same time; 3) First record the audio content in the singing interface of the target song, and then you can use the preset library (such as local In the gallery or cloud gallery), obtain the video file selected by the user and filter out the sound in the obtained video file to obtain the image content; 4) First record the audio content in the singing interface of the target song, and then Obtain the photographed picture file selected by the user from a preset gallery (such as a local gallery or a cloud gallery), and filter out the sound in the obtained picture file to obtain image content, etc.
  • a preset gallery such as a local gallery or a cloud gallery
  • the audio content can be added to the image content to generate the audio and video of the target song; it can also be the resolution, drawing, and image content of the image content. Filters and other processing, volume adjustment, playback style and other processing on the audio content, and then based on the processed audio content and image content, the audio and video of the target song is generated.
  • the audio and video generated at this time is a complete audio and video that has not undergone processing such as intercepting segments.
  • S120 If a save event is detected, save the audio and video, and determine a segment of a preset duration from the audio and video as a selected segment.
  • the save event may be triggered by the user manually or by voice, etc., and is used to request the server or the K song application provided by the server to save the audio and video events.
  • the save event may be triggered by the user manually clicking the save button in the currently displayed interface of the K song application provided by the server, where the currently displayed interface may be a preview interface of audio and video.
  • the preset duration is a preset total playback duration of the selected segment, which can be modified according to the total playback duration of audio and video in actual situations and actual needs.
  • the preset duration may be a set ratio of the total playback duration of the audio and video. For example, if the total playback duration of the audio and video is 90s and the set ratio is 1/6, the preset duration is 15s.
  • a segment with a preset duration can be randomly selected from the audio and video as the selected segment.
  • the first 15s segment in the audio and video can be selected as the selected segment by default; or it can be based on the feature information of lyrics in the audio and video. , Image feature information, etc. Select a segment of preset duration from the audio and video as the selected segment.
  • determining a clip with a preset duration from the audio and video, as a selected clip can also be: selecting a clip with a preset duration from the audio and video, and adjusting the resolution and playback speed of the selected clip to the setting Set the value to generate a selected segment. Or: generate selected fragments by interacting with users such as visualization or question and answer. For example, the selected segment can be generated by interacting with the user in the selected segment production interface provided.
  • the server Take the server as the execution subject as an example.
  • the server After the server generates the audio and video of the target song, it can send the generated audio and video to the user's karaoke application that records the audio content of the target song, so that the karaoke application displays the audio on the audio and video preview interface.
  • Video at this time, if the user is satisfied with the displayed audio and video, he can click the save button in the preview interface to save, and the server can detect the save event through the K song application, save the audio and video locally, and A segment of preset duration is determined from the audio and video as a selected segment.
  • the server may also display the generated selected segment to the user in a preview form through the user's K song application.
  • the publishing event may be triggered by the user manually or by voice, etc., and is used to request the server or the K song application provided by the server to publish selected clips or audio and video events.
  • the video list interface is used to display a list interface of selected segments of each user's audio and video.
  • the K song application provided by the server and the server have different ways of adding the selected clips to the video list interface for display after detecting the publishing event.
  • the server if the user's K song application that records the audio content of the target song detects a release event, the selected segment can be added to the video list interface, and then the server can control the multiple items managed by the server.
  • a K song application shows a video list interface for adding selected clips to users corresponding to multiple K song applications.
  • the server may add the selected clips to the video list interface, and determine the loading information of the video list interface including the selected clips; according to the loading information, load the video list interface including the selected clips and display the Featured clips, that is, the server can send the loading information to multiple K song applications managed by it, so that multiple K song applications can load the video list interface including the selected clips according to the loading information and correspond to them User impressions.
  • the loading information can be a link address, or data information of a selected fragment.
  • the selected segment of the target song can be sent to the server to request the server to add the selected segment to the video list interface and feedback loading includes The related information of the video list interface of the selected segment, and then the K song application can load and display the video list interface including the selected segment according to the relevant information; at the same time, the server can also load the video including the selected segment
  • the relevant information of the list interface is sent to the K song application of other users managed by the server, so that other users can watch the content published by the user.
  • the relevant information for loading the video list interface including the selected fragments may be the loading information of the video list interface including the selected fragments determined by the server.
  • an interface for entering audio and video is embedded in the selected fragments in this implementation.
  • the interface may be embodied in the form of a physical button or a floating ball.
  • the server detects through a K song application that the user using the K song application clicks on the physical button of the audio and video, it can send the link address or storage address of the pre-saved audio and video to the K song application, so that the K song application jumps to the complete viewing interface of the audio and video according to the obtained address and loads the audio and video for display, so that the user can watch the complete audio and video in the complete viewing interface of the audio and video .
  • a K song application detects that the user using the K song application clicks on the physical button of the audio and video, it can send a full version viewing request including the selected segment information to the server to request the server to obtain the selected segment information
  • the complete audio and video of the selected segment and feedback and then the K song application can obtain the complete audio and video feedback from the server and display it to the user, or the K song application can be based on the complete audio and video feedback from the server
  • the address jumps to the complete viewing interface of the audio and video and loads the audio and video for display.
  • a save event is detected, while saving the audio and video of the target song generated based on the recorded audio content and image content of the target song, a segment of preset duration is determined from the audio and video, As a featured clip; if a publishing event is detected later, the generated featured clip can be added to the video list interface for display.
  • this solution not only helps users quickly capture the key points and enhances user experience; it also improves the loading speed of the interface and reduces the operating power consumption of electronic devices.
  • FIG. 2 shows a flowchart of another selected fragment processing method provided by an embodiment of the present disclosure. This embodiment is described on the basis of multiple alternative solutions provided in the above embodiment. This embodiment is different from the above embodiment. In the multiple steps provided, it is introduced how to process the audio and video according to preset fragment processing rules to generate selected fragments.
  • the selected segment processing method in this embodiment may include S210 to S250.
  • S210 Generate audio and video of the target song based on the recorded audio content and image content of the target song.
  • S220 If the save event is detected, save the audio and video, and use the time when the first lyrics are sung in the audio and video or the time when the climax segment is entered as the start time of the selected segment.
  • the generated audio and video can be saved, and the first song in the audio and video can be determined based on the feature information of the lyrics in the audio and video (for example, it can include the start position and end position of the lyrics)
  • the time of the lyrics of the sentence is used as the start time of the selected segment; or the time of entering the climax segment, that is, the time of entering the first lyrics of the climax segment, can be determined as the start time of the selected segment.
  • the time when the video frame of the open mouth image is first detected in the audio and video can be determined as the start time of the selected segment.
  • S230 Determine the end time of the selected segment according to the start time and the preset duration.
  • the preset duration may be added to the start time to determine the end time of the selected segment. For example, if the start time is 0:0:5, and the preset duration is 15s, it can be determined that the end time of the selected segment is 0:0:20.
  • S240 Use the segments between the start time and the end time in the audio and video as selected segments.
  • the segment between the start time and the end time may be intercepted from the audio and video as the selected segment.
  • the technical solution provided by the embodiments of the present disclosure provides a way to determine the start time and end time of the selected segment based on the feature information of the lyrics, and then generate the selected segment based on the audio and video and the determined start time and end time; And if a publishing event is detected later, the generated selected fragments can be added to the video list interface for display.
  • this solution not only helps users quickly capture the key points and enhances user experience; it also improves the loading speed of the interface and reduces the operating power consumption of electronic devices.
  • FIG. 3 shows a flowchart of another selected fragment processing method provided by an embodiment of the present disclosure. This embodiment is described on the basis of multiple alternative solutions provided by the foregoing embodiment. This embodiment is different from the foregoing embodiment. In the multiple steps provided, it is introduced how to process the audio and video according to preset fragment processing rules to generate selected fragments.
  • the selected segment processing method in this embodiment may include S310 to S350.
  • S310 Generate audio and video of the target song according to the recorded audio content and image content of the target song.
  • S320 If the save event is detected, save the audio and video, and select a segment of the first preset duration from the audio and video as the first segment.
  • the first preset duration is a preset total playback duration of the first segment.
  • the first preset duration is less than the preset duration.
  • it may be a set ratio of the preset duration, for example, it may be 2/3 of the preset duration.
  • the generated audio and video can be saved, and a segment of the first preset duration can be randomly selected from the audio and video as the first segment.
  • a segment with the first preset duration in the audio and video can be taken as the first segment by default, or a segment with the first preset duration can be selected from the audio and video as the first segment based on the lyrics feature information and image feature information in the audio and video.
  • the first preset duration is greater than the second preset duration
  • the second segment may be an introduction segment for introducing the content of the first segment.
  • the second segment may be used to introduce the user who sang the song in the first segment this time, or to introduce the star who actually sang the song in the first segment.
  • the second segment may be a brief introduction segment, which is used to introduce the actual singing content of the first segment, etc.
  • a second preset duration segment obtains a second preset duration segment.
  • a preset gallery such as a local gallery or a cloud gallery
  • a relevant segment of a star who actually sang the song in the first segment can be obtained from a preset gallery
  • a segment of the second preset duration is intercepted from the obtained segment and used as the second segment.
  • a segment associated with the first segment uploaded by the user in real time can be acquired, and then a segment of the second preset duration can be intercepted from the acquired segment as the second segment.
  • the pictures taken by the user in real time can be acquired, and then the same accompaniment sounds as in the first segment can be acquired, and then a segment of the second preset duration can be generated as the second segment based on the pictures and accompaniment sounds taken in real time.
  • S340 Perform splicing processing on the first segment and the second segment to generate a selected segment.
  • the second segment can be placed before the first segment for splicing, and the pitch of the spliced place can be adjusted, so that it can smoothly transition from the first segment to the second segment, thereby generating a selected segment.
  • the second segment is a brief introduction segment, it can also be placed after the first segment for splicing to generate a selected segment.
  • this embodiment can generate selected segments based on complete audio and video and other acquired audio and video segments, which increases the flexibility of generating selected segments.
  • the technical solution provided by the embodiment of the present disclosure can generate a selected segment by splicing the acquired second segment and the first segment intercepted from the audio and video; afterwards, if a publishing event is detected, the generated The selected clips are added to the video list interface for display.
  • this solution not only helps users quickly capture the key points and enhances user experience; it also improves the loading speed of the interface and reduces the operating power consumption of electronic devices.
  • this solution provides a way to generate selected segments based on complete audio and video and other acquired audio and video segments, which increases the flexibility of the way to generate selected segments.
  • FIG. 4 shows a flowchart of another selected fragment processing method provided by an embodiment of the present disclosure. This embodiment is described on the basis of multiple alternative solutions provided by the foregoing embodiment. This embodiment is different from the foregoing embodiment. In the multiple steps provided, it is introduced how to process the audio and video according to preset fragment processing rules to generate selected fragments.
  • the selected segment processing method in this embodiment may include S410 to S450.
  • S410 Generate audio and video of the target song according to the recorded audio content and image content of the target song.
  • the selected segment production interface is a visual interface, which is a bridge for the server or the K song application provided by the server to interact with the user to generate the selected segment.
  • the generated audio and video can be saved, and the selected segment production interface can be shown to the user.
  • S430 Determine the start time and end time of the preset duration as the start time and end time of the selected segment according to the user's operation on the selected segment production interface.
  • the selected segment production interface can include a preview window of audio and video, audio and video progress bar, start cursor and end cursor that can be moved on the progress bar, reverberation mode adjustment options, resolution adjustment options, and volume Adjustment options, filter options, and mapping options, etc.
  • the user can click on the audio and video preview window in the selected segment production interface, and can adjust the start cursor and end cursor in the audio and video according to the audio and video played in real time, and can also display the audio and video according to actual needs.
  • a clip is added with drawings, etc.; and then the K song application provided by the server or the server can adjust the start cursor and end cursor in the audio and video on the selected clip production interface to determine the start of the preset duration Time and end time, and use the preset start time and end time as the start time and end time of the selected clip.
  • the position where the end cursor can be adjusted can be dynamically displayed, so as to increase the interest of the user during the operation.
  • the preset duration can be automatically modified according to the user's operation to meet the user's requirements. demand.
  • the segment between the start time and the end time may be intercepted from the audio and video as the selected segment. Or, if the user adds a drawing operation to the segment between the start time and the end time, you can also first intercept the segment between the start time and the end time from the audio and video, and then map the intercepted segment Paper and other processing to generate selected fragments.
  • the technical solution provided by the embodiments of the present disclosure shows the user a selected segment production interface, and determines the start time and end time of the preset duration as the start time of the selected segment according to the user's operation on the selected production interface And the end time, and then generate a selected segment; if a publishing event is detected later, the generated selected segment can be added to the video list interface for display.
  • this solution not only helps users quickly capture the key points and enhances user experience; it also improves the loading speed of the interface and reduces the operating power consumption of electronic devices.
  • this solution provides a way to generate selected fragments based on visualization, so that the generated selected fragments are more in line with the needs of users.
  • FIG. 5 shows a schematic structural diagram of a selected segment processing device provided by an embodiment of the present disclosure.
  • the embodiment of the present disclosure may be applicable to how to generate and display the selected segment in the video list interface, so as to solve the problem that related technologies directly integrate the complete
  • the display of audio and video in the video list interface is not convenient for the user to quickly capture the key points, which leads to a poor user experience.
  • the device can be implemented by software and/or hardware, and can be configured on an electronic device.
  • the so-called electronic device may be a server device that carries the function of processing selected fragments, or a terminal device that configures the K song application provided by the server.
  • the selected segment processing device in the embodiment of the present disclosure includes: an audio and video generation module 510, a selected segment generation module 520, and an add display module 530.
  • the audio and video generation module 510 is configured to generate the audio and video of the target song according to the recorded audio content and image content of the target song;
  • the selected segment generating module 520 is configured to save the audio and video if a saving event is detected, and determine a segment with a preset duration from the audio and video as the selected segment;
  • the display adding module 530 is configured to add the selected segment to the video list interface for display if a publishing event is detected.
  • the selected fragment generation module 520 is set to:
  • the segments between the start time and the end time in the audio and video are regarded as selected segments.
  • the selected fragment generation module 520 is still set to:
  • the splicing process is performed on the first segment and the second segment to generate a selected segment.
  • the second segment is an introduction segment, which is used to introduce the content of the first segment.
  • the selected fragment generation module 520 is also set to:
  • the selection segment production interface determine the start time and end time of the preset duration as the start time and end time of the selection segment
  • the segments between the start time and the end time in the audio and video are regarded as selected segments.
  • an interface for entering audio and video is embedded in the selected segment, and the interface is embodied in the form of a physical button or a floating ball.
  • the adding display module 530 is set to:
  • FIG. 6 shows a schematic structural diagram of an electronic device 600 suitable for implementing the embodiments of the present disclosure.
  • the electronic devices in the embodiments of the present disclosure may include, for example, mobile phones, notebook computers, digital broadcast receivers, personal digital assistants (Personal Digital Assistant, PDA), tablet computers (Portable Android Device, PAD), and portable multimedia players (Personal Multimedia). Player, PMP), mobile terminals such as in-vehicle terminals (for example, in-vehicle navigation terminals), and fixed terminals such as digital televisions (television, TV), desktop computers, etc.
  • PDA Personal Digital Assistant
  • PAD Portable Android Device
  • PMP portable multimedia players
  • PMP mobile terminals
  • in-vehicle terminals for example, in-vehicle navigation terminals
  • fixed terminals such as digital televisions (television, TV), desktop computers, etc.
  • the so-called electronic device in this embodiment may be a server device that carries the function of processing selected fragments, or a terminal device that configures a K song application provided by the server.
  • the electronic device shown in FIG. 6 is only an example, and should not bring any limitation to the function and scope of use of the embodiments of the present disclosure.
  • the electronic device 600 may include a processing device (such as a central processing unit, a graphics processor, etc.) 601.
  • the processing device 601 may be based on a program stored in a read-only memory (Read-only Memory, ROM) 602 or from a storage device.
  • the device 608 loads a program in a random access memory (Random Access Memory, RAM) 603 to perform various appropriate actions and processes.
  • the RAM 603 also stores various programs and data required for the operation of the electronic device 600.
  • the processing device 601, the ROM 602, and the RAM 603 are connected to each other through a bus 604.
  • An input/output (Input/Output, I/O) interface 605 is also connected to the bus 604.
  • the following devices can be connected to the I/O interface 605: including input devices 606 such as touch screen, touch pad, keyboard, mouse, camera, microphone, accelerometer, gyroscope, etc.; including, for example, a liquid crystal display (Liquid Crystal Display, LCD), speaker, vibrator, etc. output device 607; storage device 608 including, for example, magnetic tape, hard disk, etc.; and communication device 609.
  • the communication device 609 may allow the electronic device 600 to perform wireless or wired communication with other devices to exchange data.
  • FIG. 6 shows an electronic device 600 with multiple devices, it is not required to implement or have all the illustrated devices, and may alternatively implement or have more or fewer devices.
  • the process described above with reference to the flowchart may be implemented as a computer software program.
  • the embodiments of the present disclosure include a computer program product, which includes a computer program carried on a computer-readable medium, and the computer program contains program code for executing the method shown in the flowchart.
  • the computer program may be downloaded and installed from the network through the communication device 609, or installed from the storage device 608, or installed from the ROM 602.
  • the processing device 601 the above-mentioned functions defined in the method of the embodiment of the present disclosure are executed.
  • the aforementioned computer-readable medium of the present disclosure may be a computer-readable signal medium or a computer-readable storage medium or any combination of the two.
  • the computer-readable storage medium may be, for example, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or a combination of any of the above.
  • Computer-readable storage media may include: electrical connections with one or more wires, portable computer disks, hard disks, RAM, ROM, Erasable Programmable Read-Only Memory (EPROM) or flash memory, optical fiber , Portable Compact Disc Read-Only Memory (CD-ROM), optical storage device, magnetic storage device, or any suitable combination of the above.
  • a computer-readable storage medium may be any tangible medium that contains or stores a program, and the program may be used by or in combination with an instruction execution system, apparatus, or device.
  • a computer-readable signal medium may include a data signal propagated in a baseband or as a part of a carrier wave, and a computer-readable program code is carried therein. This propagated data signal can take many forms, including electromagnetic signals, optical signals, or any suitable combination of the above.
  • the computer-readable signal medium may also be any computer-readable medium other than the computer-readable storage medium.
  • the computer-readable signal medium may send, propagate, or transmit the program for use by or in combination with the instruction execution system, apparatus, or device .
  • the program code contained on the computer-readable medium can be transmitted by any suitable medium, including: wire, optical cable, radio frequency (RF), etc., or any suitable combination of the above.
  • the above-mentioned computer-readable medium may be included in the above-mentioned terminal or server; or it may exist alone without being installed in the server.
  • the above-mentioned computer-readable medium carries one or more programs.
  • the server is caused to generate the audio and video of the target song according to the recorded audio content and image content of the target song ; If a save event is detected, the audio and video will be saved, and the audio and video will be processed according to the preset segment processing rules to generate a selected segment; if a publishing event is detected, the selected segment will be added to the video list interface for display .
  • the computer program code used to perform the operations of the present disclosure may be written in one or more programming languages or a combination thereof.
  • the above-mentioned programming languages include object-oriented programming languages—such as Java, Smalltalk, C++, and also conventional Procedural programming language-such as "C" language or similar programming language.
  • the program code can be executed entirely on the user's computer, partly on the user's computer, executed as an independent software package, partly on the user's computer and partly executed on a remote computer, or entirely executed on the remote computer or server.
  • the remote computer can be connected to the user's computer through any kind of network-including Local Area Network (LAN) or Wide Area Network (WAN)-or it can be connected to an external computer (for example, use an Internet service provider to connect via the Internet).
  • LAN Local Area Network
  • WAN Wide Area Network
  • each block in the flowchart or block diagram can represent a module, program segment, or part of code, and the module, program segment, or part of code contains one or more for realizing the specified logical function Executable instructions.
  • the functions marked in the block may also occur in a different order from the order marked in the drawings. For example, two blocks shown in succession can actually be executed substantially in parallel, or they can sometimes be executed in the reverse order, depending on the functions involved.
  • Each block in the block diagram and/or flowchart, and the combination of the blocks in the block diagram and/or flowchart, can be implemented by a dedicated hardware-based system that performs the specified functions or operations, or can be implemented by dedicated hardware Realized in combination with computer instructions.
  • the units described in the embodiments of the present disclosure can be implemented in software or hardware. Among them, the name of the unit in one case does not constitute a limitation on the unit itself.
  • exemplary types of hardware logic components include: Field Programmable Gate Array (FPGA), Application Specific Integrated Circuit (ASIC), and application specific standard products (Application Specific Standard Parts, ASSP), System-on-a-Chip (SOC), Complex Programmable Logic Device (CPLD), etc.
  • FPGA Field Programmable Gate Array
  • ASIC Application Specific Integrated Circuit
  • ASSP Application Specific Standard Parts
  • SOC System-on-a-Chip
  • CPLD Complex Programmable Logic Device
  • a machine-readable medium may be a tangible medium, and the machine-readable medium may contain or store a program for use by the instruction execution system, apparatus, or device or in combination with the instruction execution system, apparatus, or device.
  • the machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium.
  • the machine-readable medium may include an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, device, or device, or any suitable combination of the foregoing.
  • Machine-readable storage media include electrical connections based on one or more lines, portable computer disks, hard disks, RAM, ROM, EPROM or flash memory, optical fibers, CD-ROMs, optical storage devices, magnetic storage devices, or the foregoing Any suitable combination.

Landscapes

  • Engineering & Computer Science (AREA)
  • Multimedia (AREA)
  • Signal Processing (AREA)
  • Human Computer Interaction (AREA)
  • Databases & Information Systems (AREA)
  • Computer Security & Cryptography (AREA)
  • User Interface Of Digital Computer (AREA)
  • Two-Way Televisions, Distribution Of Moving Picture Or The Like (AREA)

Abstract

本实施例公开了一种精选片段处理方法、装置、电子设备及可读介质。其中,该精选片段处理方法包括:依据录制的目标歌曲的音频内容和图像内容,生成目标歌曲的音视频;在检测到保存事件的情况下,保存所述音视频,并从所述音视频中确定一个预设时长的片段,作为精选片段;在检测到发布事件的情况下,将所述精选片段添加到视频列表界面中进行展示。

Description

精选片段处理方法、装置、电子设备及可读介质
本申请要求在2019年06月27日提交中国专利局、申请号为201910569485.9的中国专利申请的优先权,该申请的全部内容通过引用结合在本申请中。
技术领域
本公开实施例涉及互联网技术领域,例如涉及一种精选片段处理方法、装置、电子设备及可读介质。
背景技术
在K歌类应用程序中,用户可以观看其他用户发布的音视频,也可以选择自己喜爱的歌曲录制音视频并发布。
用户可以在搜索框中输入喜爱的歌曲名,而后点击搜索结果中任一K歌选项,进入该歌曲的演唱界面,进行歌曲的录制。而且通常情况下,在歌曲录制完成之后,用户可以选择录制视频图像,进而可根据已录制的歌曲的音频和视频图像生成并保存音视频,并在检测到用户的发布指令之后,直接将所保存的音视频发布在视频列表界面,以便其他用户观看。
然而,在用户所发布的音视频较长的情况下,不便于观看该音视频的用户快速捕捉到重点,且容易出现情绪化现象,不利用于用户体验。
发明内容
本公开实施例提供了一种精选片段处理方法、装置、电子设备及可读介质,有助于用户快速捕捉到重点,进而提升了用户体验。
本公开实施例提供了一种精选片段处理方法,该方法包括:
依据录制的目标歌曲的音频内容和图像内容,生成目标歌曲的音视频;
在检测到保存事件的情况下,保存所述音视频,并从所述音视频中确定一个预设时长的片段,作为精选片段;
在检测到发布事件的情况下,将所述精选片段添加到视频列表界面中进行展示。
本公开实施例还提供了一种精选片段处理装置,该装置包括:
音视频生成模块,设置为依据录制的目标歌曲的音频内容和图像内容,生成目标歌曲的音视频;
精选片段生成模块,设置为在检测到保存事件的情况下,保存所述音视频,并从所述音视频中确定一个预设时长的片段,作为精选片段;
添加展示模块,设置为在检测到发布事件的情况下,将所述精选片段添加到视频列表界面中进行展示。
本公开实施例还提供了一种电子设备,该电子设备包括:
一个或多个处理器;
存储器,设置为存储一个或多个程序;
当所述一个或多个程序被所述一个或多个处理器执行,使得所述一个或多个处理器实现如本公开任意实施例所述的精选片段处理方法。
本公开实施例提供了一种可读介质,可读介质上存储有计算机程序,所述计算机程序被处理器执行时实现如本公开任意实施例所述的精选片段处理方法。
附图说明
图1示出了本公开实施例提供的一种精选片段处理方法的流程图;
图2示出了本公开实施例提供的另一种精选片段处理方法的流程图;
图3示出了本公开实施例提供的另一种精选片段处理方法的流程图;
图4示出了本公开实施例提供的另一种精选片段处理方法的流程图;
图5示出了本公开实施例提供的一种精选片段处理装置的结构示意图;
图6示出了本公开实施例提供的一种电子设备的结构示意图。
具体实施方式
下面将参照附图描述本公开的实施例。虽然附图中显示了本公开的一些实施例,本公开可以通过多种形式来实现,而且不应该被解释为限于这里阐述的实施例。
本公开的方法实施方式中记载的多个步骤可以按照不同的顺序执行,和/或并行执行。此外,方法实施方式可以包括附加的步骤和/或省略执行示出的步骤。本公开的范围在此方面不受限制。下述多个实施例中,每个实施例中同时提供了可选特征和示例,实施例中记载的多个特征可进行组合,形成多个可选方案, 不应将每个编号的实施例仅视为一个技术方案。
本文使用的术语“包括”及其变形是开放性包括,即“包括但不限于”。术语“基于”是“至少部分地基于”。术语“一个实施例”表示“至少一个实施例”;术语“另一实施例”表示“至少一个另外的实施例”;术语“一些实施例”表示“至少一些实施例”。其他术语的相关定义将在下文描述中给出。
本公开中提及的“第一”、“第二”等概念仅用于对不同的装置、模块或单元进行区分,并非用于限定这些装置、模块或单元所执行的功能的顺序或者相互依存关系。
本公开中提及的“一个”、“多个”的修饰是示意性而非限制性的,除非在上下文另有明确指出,否则应该理解为“一个或多个”。
图1示出了本公开实施例提供的一种精选片段处理方法的流程图,本公开实施例可适用于如何生成并在视频列表界面中展示精选片段,以解决相关技术直接将完整的音视频展示在视频列表界面中不便于用户快速捕捉到重点,进而导致用户体验不佳的情况。该方法可以由本公开实施例提供的精选片段处理装置或电子设备来执行,该装置可以通过软件和/或硬件的方式来实现。可选的,所谓电子设备可以是承载精选片段处理功能的服务端设备,还可以是配置服务端所提供的K歌应用程序的终端设备等。
可选的,如图1所示,本公开实施例中提供的精选片段处理方法包括S110至S130。
S110、依据录制的目标歌曲的音频内容和图像内容,生成目标歌曲的音视频。
本实施例中,目标歌曲是指用户所选择录制的音频和/或图像的歌曲。可选的,可以采用下述任一种方式录制目标歌曲的音频内容和图像内容:1)先在目标歌曲的演唱界面中录制音频内容,而后将已录制的音频内容作为视频原声,在视频录制界面中,录制图像内容;2)在音视频同步录制界面,同时录制目标歌曲的音频内容和图像内容;3)先在目标歌曲的演唱界面中录制音频内容,而后可从预设图库(如本地图库或云端图库)中获取用户选择的已拍摄好的视频文件,并滤除所获取的视频文件中的声音,以得到图像内容;4)先在目标歌曲的演唱界面中录制音频内容,而后可从预设图库(如本地图库或云端图库)中获取用户选择的已拍摄好的图片文件,并滤除所获取的图片文件中的声音,以得到图像内容等。
在一实施例中,在录制目标歌曲的音频内容和图像内容之后,可以将音频 内容添加到图像内容中,以生成目标歌曲的音视频;还可以是先对图像内容进行分辨率、贴图纸以及滤镜等处理,对音频内容进行音量调节、播放风格等处理,而后依据处理的音频内容和图像内容,生成目标歌曲的音视频等。此外,还可以依据目标歌曲的音频内容和图像内容,采用其他可实施方式生成目标歌曲的音视频,本实施例对此不做限定。在一实施例中,此时生成的音视频是未经过截取片段等处理的完整的音视频。
S120、若检测到保存事件,则保存音视频,并从音视频中确定一个预设时长的片段,作为精选片段。
本实施例中,保存事件可以是用户手动或语音等形式触发产生的,用于请求服务端或服务端所提供的K歌应用程序保存音视频的事件。例如,保存事件可以是用户手动点击服务端所提供的K歌应用程序当前所展示的界面中的保存按钮所触发产生的,其中当前所展示的界面可以是音视频的预览界面等。
在一实施例中,预设时长是预先设定的精选片段的播放总时长,可根据实际情况中音视频的播放总时长以及实际需求等进行修正。在一实施例中,预设时长可以是音视频的播放总时长的设定比例,如音视频的播放总时长为90s,设定比例为1/6,则预设时长为15s。
在一实施例中,可随机从音视频中选择一个预设时长的片段作为精选片段,如可以默认将音视频中前15s的片段作为精选片段;或者还可以依据音视频中歌词特征信息、图像特征信息等从音视频中选择一个预设时长的片段作为精选片段。
示例性的,从音视频中确定一个预设时长的片段,作为精选片段还可以是:从音视频中选择一个预设时长的片段,并将所选择片段的分辨率、播放速度调整至设定数值,以生成精选片段。或者是:通过与用户进行可视化或问答等交互的方式,生成精选片段。例如,可通过与用户在提供的精选片段制作界面进行交互,进行生成精选片段。
以服务端是执行主体为例进行说明。服务端在生成目标歌曲的音视频之后,可以将生成的音视频发送至录制目标歌曲的音频内容等的用户的K歌应用程序,以使该K歌应用程序在音视频的预览界面展示该音视频;此时,若该用户对所展示的音视频满意,可以点击该预览界面中保存按钮进行保存,进而服务端可通过该K歌应用程序检测到保存事件,将音视频保存在本地,并从音视频中确定一个预设时长的片段,作为精选片段。在一实施例中,在生成精选片段之后,服务端还可以通过该用户的K歌应用程序将生成的精选片段以预览形式展示给用户。
此外,为了便于满足用户的不同需求,示例性的,在从音视频中确定一个预设时长的片段,作为精选片段之前,还可以包括对音视频进行备份保存,以便后续用户不仅可观看精选片段,还可以观看完整的音视频。
S130、若检测到发布事件,则将精选片段添加到视频列表界面中进行展示。
本实施例中,发布事件可以是用户手动或语音等形式触发产生的,用于请求服务端或服务端所提供的K歌应用程序发布精选片段或音视频的事件。视频列表界面用于展示每个用户的音视频的精选片段的列表界面。
可选的,服务端和服务端所提供的K歌应用程序在检测到发布事件后,将精选片段添加到视频列表界面中进行展示的方式不同。例如,对于服务端而言,若通过录制目标歌曲的音频内容等的用户的K歌应用程序检测到发布事件,则可以将精选片段添加至视频列表界面,而后可控制服务端所管理的多个K歌应用程序向多个K歌应用程序对应的用户展示添加精选片段的视频列表界面。在一实施例中,服务端可以将精选片段添加至视频列表界面,并确定包括精选片段的视频列表界面的加载信息;依据加载信息,加载包括精选片段的视频列表界面并展示所述精选片段,即服务端可向其所管理的多个K歌应用程序发送该加载信息,以使多个K歌应用程序依据该加载信息,加载包括精选片段的视频列表界面并向其对应的用户展示。其中,加载信息可以是链接地址,也可以是精选片段的数据信息等。
对于服务端所提供的K歌应用程序而言,若检测到发布事件,则可以将目标歌曲的精选片段发送至服务端,以请求服务端将精选片段添加至视频列表界面并反馈加载包括精选片段的视频列表界面的相关信息,进而该K歌应用程序可依据该相关信息加载包括精选片段的视频列表界面并展示;与此同时,服务端还可将加载包括精选片段的视频列表界面的相关信息发送至服务端所管理的其他用户的K歌应用程序,以便其他用户可观看到该用户发布的内容。其中加载包括精选片段的视频列表界面的相关信息可以是服务端所确定的包括精选片段的视频列表界面的加载信息。
由于实际场景中用户可能具有观看完整的音视频的需求,可选的,本实施了中精选片段中嵌入有进入音视频的接口。在一实施例中,该接口可以以物理按钮或悬浮球等形式体现。
例如,在播放精选片段的过程中,当检测到用户点击音视频的物理按钮,则可以向用户展示该音视频的完整观看界面,以便用户可在该完整观看界面中观看完整的音视频。在一实施例中,服务端若通过一个K歌应用程序检测到使用该K歌应用程序的用户点击音视频的物理按钮,则可以将预先保存的音视频的链接地址或存储地址等发送至该K歌应用程序,以便该K歌应用程序依据所 获取的地址跳转至该音视频的完整观看界面并加载音视频进行展示,以便用户可在该音视频的完整观看界面中观看完整的音视频。若一个K歌应用程序检测到使用该K歌应用程序的用户点击音视频的物理按钮,则可以向服务端发送包括精选片段信息的完整版观看请求,以请求服务端依据精选片段信息获取该精选片段的完整的音视频并反馈,进而该K歌应用程序可获取服务端反馈的完整的音视频并展示给用户,或者该K歌应用程序可依据服务端反馈的完整的音视频的地址跳转至该音视频的完整观看界面并加载音视频进行展示。
本实施例中,通过在视频列表界面中展示精选片段,不仅可以有助于用户快速捕捉到重点,增加用户观看的兴趣,提升了用户体验;而且还可提升界面的加载速度,以及降低电子设备的运行功耗等。此外,通过在精选片段中嵌入进入完整的音视频的接口,又可满足用户想要观看完整的音视频的需求。
本公开实施例提供的技术方案,若检测到保存事件,在保存依据录制的目标歌曲的音频内容和图像内容生成的目标歌曲的音视频的同时,从音视频中确定一个预设时长的片段,作为精选片段;之后若检测到发布事件,则可以将所生成的精选片段添加到视频列表界面中进行展示。本方案通过在视频列表界面中展示精选片段,不仅有助于用户快速捕捉到重点,提升用户体验;而且还可提升界面的加载速度,以及降低电子设备的运行功耗等。
图2示出了本公开实施例提供的另一种精选片段处理方法的流程图,本实施例在上述实施例提供的多个可选方案的基础上进行描述,本实施例对于上述实施例提供的多个步骤中如何根据预设的片段处理规则对所述音视频进行处理,生成精选片段进行介绍。
可选的,如图2所示,本实施例中的精选片段处理方法可以包括S210至S250。
S210、依据录制的目标歌曲的音频内容和图像内容,生成目标歌曲的音视频。
S220、若检测到保存事件,则保存音视频,并将音视频中演唱第一句歌词的时间,或进入高潮片段的时间作为精选片段的起始时间。
本实施例中,若检测到保存事件,则可以保存所生成的音视频,并可以依据音视频中歌词特征信息(如可以包括歌词起始位置、结束位置等),确定音视频中演唱第一句歌词的时间作为精选片段的起始时间;或者可以确定进入高潮片段的时间,即进入高潮片段的第一句歌词的时间,作为精选片段的起始时间。
此外,还可以依据音视频中图像特征信息如嘴部特征信息,确定音视频中首次检测到张嘴图像视频帧的时间,作为精选片段的起始时间。
S230、依据起始时间以及预设时长,确定精选片段的结束时间。
在一实施例中,可以在起始时间的基础上累加上预设时长,即可确定精选片段的结束时间。例如,起始时间为0:0:5,预设时长为15s,则可以确定精选片段的结束时间为0:0:20。
S240、将音视频中起始时间和结束时间之间的片段,作为精选片段。
在一实施例中,在确定精选片段的起始时间和结束时间之后,可以从音视频中截取起始时间和结束时间之间的片段,作为精选片段。或者,还可以先从音视频中截取起始时间和结束时间之间的片段,而后对所截取的片段进行贴图纸等处理,生成精选片段。
S250、若检测到发布事件,则将精选片段添加到视频列表界面中进行展示。
本公开实施例提供的技术方案,提供了一种基于歌词特征信息确定精选片段的起始时间和结束时间,进而依据音视频以及所确定的起始时间和结束时间生成精选片段的方式;而且之后若检测到发布事件,则可以将所生成的精选片段添加到视频列表界面中进行展示。本方案通过在视频列表界面中展示精选片段,不仅有助于用户快速捕捉到重点,提升用户体验;而且还可提升界面的加载速度,以及降低电子设备的运行功耗等。
图3示出了本公开实施例提供的另一种精选片段处理方法的流程图,本实施例在上述实施例提供的多个可选方案的基础上进行描述,本实施例对于上述实施例提供的多个步骤中如何根据预设的片段处理规则对所述音视频进行处理,生成精选片段进行介绍。
可选的,如图3所示,本实施例中的精选片段处理方法可以包括S310至S350。
S310、依据录制的目标歌曲的音频内容和图像内容,生成目标歌曲的音视频。
S320、若检测到保存事件,则保存音视频,并从音视频中选择一个第一预设时长的片段,作为第一片段。
本实施例中,第一预设时长是预先设定的第一片段的播放总时长。可选的,第一预设时长小于预设时长,在一实施例中,可以是预设时长的设定比例,如可以是设定时长的2/3等。
在一实施例中,若检测到保存事件,则可以保存所生成的音视频,并可以随机从音视频中选择一个第一预设时长的片段作为第一片段。如可以默认将音视频中前第一预设时长的片段作为第一片段,或者还可以依据音视频中歌词特征信息、图像特征信息等从音视频中选择一个第一预设时长的片段作为第一片段。
S330、获取一个第二预设时长的片段,作为第二片段。
本实施例中,第一预设时长大于第二预设时长,且第二片段可以为介绍片段,用于对第一片段的内容进行介绍。在一实施例中,第二片段可以用于介绍此次演唱第一片段中歌曲的用户,或者用于介绍实际演唱第一片段中歌曲的明星等。此外,第二片段可以为概要简介片段,用于介绍第一片段实质演唱的内容等。
可选的,获取一个第二预设时长的片段,作为第二片段可以采用下述任一种方式:1)从预设图库(如本地图库或云端图库)中获取一个与第一片段相关联的片段(例如可以从预设图库中获取实际演唱第一片段中歌曲的明星的相关片段),而后从所获取的片段中截取第二预设时长的片段,作为第二片段。2)可获取用户实时上传的与第一片段相关联的片段,而后从所获取的片段中截取第二预设时长的片段,作为第二片段。3)可以获取用户实时拍摄的图片,而后可以获取与第一片段中相同的伴奏音,进而可依据实时拍摄的图片和伴奏音,生成一个第二预设时长的片段作为第二片段等。
S340、对第一片段和第二片段进行拼接处理,生成精选片段。
可选的,可以将第二片段放置于第一片段之前进行拼接,并可调整拼接之处的音调等,使之可从第一片段平滑过渡到第二片段,进而生成精选片段。此外,若第二片段为概要简介片段,也可置于第一片段之后进行拼接,生成精选片段等。
在一实施例中,本实施例可基于完整的音视频,以及获取的其他音视频片段生成精选片段,增加了生成精选片段的灵活度。
S350、若检测到发布事件,则将精选片段添加到视频列表界面中进行展示。
本公开实施例提供的技术方案,通过对获取的第二片段,以及从音视频中截取的第一片段进行拼接处理,可以生成精选片段;之后若检测到发布事件,则可以将所生成的精选片段添加到视频列表界面中进行展示。本方案通过在视频列表界面中展示精选片段,不仅有助于用户快速捕捉到重点,提升用户体验;而且还可提升界面的加载速度,以及降低电子设备的运行功耗等。此外,本方案提供了一种可基于完整的音视频,以及获取的其他音视频片段生成精选片段 的方式,增加了生成精选片段方式的灵活度。
图4示出了本公开实施例提供的另一种精选片段处理方法的流程图,本实施例在上述实施例提供的多个可选方案的基础上进行描述,本实施例对于上述实施例提供的多个步骤中如何根据预设的片段处理规则对所述音视频进行处理,生成精选片段进行介绍。
可选的,如图4所示,本实施例中的精选片段处理方法可以包括S410至S450。
S410、依据录制的目标歌曲的音频内容和图像内容,生成目标歌曲的音视频。
S420、若检测到保存事件,则保存音视频,并向用户展示精选片段制作界面。
本实施例中,精选片段制作界面是一种可视化界面,是服务端或服务端所提供的K歌应用程序与用户进行交互,生成精选片段的桥梁。
在一实施例中,若检测到保存事件,则可以保存所生成的音视频,并可以向用户展示精选片段制作界面。
S430、依据用户在精选片段制作界面上的操作,确定预设时长的起始时间和结束时间作为精选片段的起始时间和结束时间。
可选的,精选片段制作界面中可包括音视频的预览窗口、音视频进度条、可在该进度条上移动的起始游标和结束游标、混响模式调节选项、分辨率调节选项、音量调节选项、滤镜选项以及贴图纸选项等。
在一实施例中,用户可以点击精选片段制作界面中音视频的预览窗口,并可依据实时播放的音视频调整音视频中的起始游标和结束游标,还可以根据实际需求在音视频中一个片段增加贴图纸等;进而服务端或服务端所提供的K歌应用程序可基于用户在精选片段制作界面上调整音视频中起始游标和结束游标的操作,确定预设时长的起始时间和结束时间,并将预设时长的起始时间和结束时间作为精选片段的起始时间和结束时间。
此外,在一实施例中,若检测到用户在调整音视频中的起始游标和结束游标的过程中,出现起始游标和结束游标之间的片段时长大于预设时长时,还可以以语音或文字等形式提醒用户精选片段的播放总时长不能超过预设时长如15s等。同时,在用户调整了起始游标之后,可以动态的显示可将结束游标调整的位置等,以增加用户操作过程中的趣味性。
同时,若检测到用户连续多次数所调整的音视频中的起始游标和结束游标之间的片段时长均小于预设时长,则可以依据用户的操作,自动修改预设时长,以满足用户的需求。
S440、将音视频中起始时间和结束时间之间的片段,作为精选片段。
在一实施例中,在确定精选片段的起始时间和结束时间之后,可以从音视频中截取起始时间和结束时间之间的片段,作为精选片段。或者,若用户在起始时间和结束时间之间的片段中增加了贴图纸操作,则还可以先从音视频中截取起始时间和结束时间之间的片段,而后对所截取的片段进行贴图纸等处理,生成精选片段。
S450、若检测到发布事件,则将精选片段添加到视频列表界面中进行展示。
本公开实施例提供的技术方案,通过向用户展示精选片段制作界面,并依据用户在精选制作界面上的操作,确定预设时长的起始时间和结束时间作为精选片段的起始时间和结束时间,进而生成精选片段;之后若检测到发布事件,则可以将所生成的精选片段添加到视频列表界面中进行展示。本方案通过在视频列表界面中展示精选片段,不仅有助于用户快速捕捉到重点,提升用户体验;而且还可提升界面的加载速度,以及降低电子设备的运行功耗等。此外,本方案提供了一种基于可视化方式生成精选片段,使得所生成的精选片段更加符合用户的需求。
图5示出了本公开实施例提供的一种精选片段处理装置的结构示意图,本公开实施例可适用于如何生成并在视频列表界面中展示精选片段,以解决相关技术直接将完整的音视频展示在视频列表界面中不便于用户快速捕捉到重点,进而导致用户体验不佳的情况,该装置可以通过软件和/或硬件来实现,可以配置于电子设备上。可选的,所谓电子设备可以是承载精选片段处理功能的服务端设备,还可以是配置服务端所提供的K歌应用程序的终端设备等。如图5所示,本公开实施例中精选片段处理装置,包括:音视频生成模块510、精选片段生成模块520和添加展示模块530。
音视频生成模块510设置为依据录制的目标歌曲的音频内容和图像内容,生成目标歌曲的音视频;
精选片段生成模块520设置为若检测到保存事件,保存所述音视频,并从音视频中确定一个预设时长的片段,作为精选片段;
添加展示模块530设置为若检测到发布事件,则将精选片段添加到视频列表界面中进行展示。
示例性的,精选片段生成模块520是设置为:
将音视频中演唱第一句歌词的时间,或进入高潮片段的时间作为精选片段的起始时间;
依据起始时间以及预设时长,确定精选片段的结束时间;
将音视频中起始时间和结束时间之间的片段,作为精选片段。
示例性的,精选片段生成模块520还是设置为:
从音视频中选择一个第一预设时长的片段,作为第一片段;
获取一个第二预设时长的片段,作为第二片段;其中第一预设时长大于第二预设时长;
对第一片段和第二片段进行拼接处理,生成精选片段。
示例性的,第二片段为介绍片段,用于对第一片段的内容进行介绍。
示例性的,精选片段生成模块520还设置为:
向用户展示精选片段制作界面;
依据用户在精选片段制作界面上的操作,确定预设时长的起始时间和结束时间作为精选片段的起始时间和结束时间;
将音视频中起始时间和结束时间之间的片段,作为精选片段。
示例性的,本实施例中精选片段中嵌入有进入音视频的接口,接口以物理按钮或悬浮球形式体现。
示例性的,添加展示模块530是设置为:
将精选片段添加至视频列表界面,并确定包括精选片段的视频列表界面的加载信息;
依据加载信息,加载包括精选片段的视频列表界面并展示所述精选片段。
本公开实施例提供的精选片段处理装置未在本公开实施例中描述的技术细节可参见上述实施例,并且本公开实施例与上述实施例具有相同的有益效果。
参见图6,图6示出了适于用来实现本公开实施例的电子设备600的结构示意图。本公开实施例中的电子设备可以包括诸如移动电话、笔记本电脑、数字广播接收器、个人数字助理(Personal Digital Assistant,PDA)、平板电脑(Portable Android Device,PAD)、便携式多媒体播放器(Personal Multimedia Player,PMP)、车载终端(例如车载导航终端)等等的移动终端以及诸如数字电视(television, TV)、台式计算机等等的固定终端。可选的,本实施例中所谓电子设备可以是承载精选片段处理功能的服务端设备,还可以是配置服务端所提供的K歌应用程序的终端设备等。图6示出的电子设备仅仅是一个示例,不应对本公开实施例的功能和使用范围带来任何限制。
如图6所示电子设备600可以包括处理装置(例如中央处理器、图形处理器等)601,处理装置601可以根据存储在只读存储器(Read-only Memory,ROM)602中的程序或者从存储装置608加载到随机访问存储器(Random Access Memory,RAM)603中的程序而执行多种适当的动作和处理。在RAM 603中,还存储有电子设备600操作所需的多种程序和数据。处理装置601、ROM 602以及RAM 603通过总线604彼此相连。输入/输出(Input/Output,I/O)接口605也连接至总线604。
在一实施例中,以下装置可以连接至I/O接口605:包括例如触摸屏、触摸板、键盘、鼠标、摄像头、麦克风、加速度计、陀螺仪等的输入装置606;包括例如液晶显示器(Liquid Crystal Display,LCD)、扬声器、振动器等的输出装置607;包括例如磁带、硬盘等的存储装置608;以及通信装置609。通信装置609可以允许电子设备600与其他设备进行无线或有线通信以交换数据。虽然图6示出了具有多种装置的电子设备600,但是,并不要求实施或具备所有示出的装置,可以替代地实施或具备更多或更少的装置。
根据本公开的实施例,上文参考流程图描述的过程可以被实现为计算机软件程序。例如,本公开的实施例包括一种计算机程序产品,其包括承载在计算机可读介质上的计算机程序,该计算机程序包含用于执行流程图所示的方法的程序代码。在这样的实施例中,该计算机程序可以通过通信装置609从网络上被下载和安装,或者从存储装置608被安装,或者从ROM 602被安装。在该计算机程序被处理装置601执行时,执行本公开实施例的方法中限定的上述功能。
在一实施例中,本公开上述的计算机可读介质可以是计算机可读信号介质或者计算机可读存储介质或者是上述两者的任意组合。计算机可读存储介质例如可以是电、磁、光、电磁、红外线、或半导体的系统、装置或器件,或者任意以上的组合。计算机可读存储介质可以包括:具有一个或多个导线的电连接、便携式计算机磁盘、硬盘、RAM、ROM、可擦式可编程只读存储器(Erasable Programmable Read-Only Memory,EPROM)或闪存、光纤、便携式紧凑磁盘只读存储器(Compact Disc Read-Only Memory,CD-ROM)、光存储器件、磁存储器件、或者上述的任意合适的组合。在本公开中,计算机可读存储介质可以是任何包含或存储程序的有形介质,该程序可以被指令执行系统、装置或者器件使用或者与其结合使用。而在本公开中,计算机可读信号介质可以包括在基 带中或者作为载波一部分传播的数据信号,其中承载了计算机可读的程序代码。这种传播的数据信号可以采用多种形式,包括电磁信号、光信号或上述的任意合适的组合。计算机可读信号介质还可以是计算机可读存储介质以外的任何计算机可读介质,该计算机可读信号介质可以发送、传播或者传输用于由指令执行系统、装置或者器件使用或者与其结合使用的程序。计算机可读介质上包含的程序代码可以用任何适当的介质传输,包括:电线、光缆、射频(Radio Frequency,RF)等等,或者上述的任意合适的组合。
上述计算机可读介质可以是上述终端或服务器中所包含的;也可以是单独存在,而未装配入该服务器中。
上述计算机可读介质承载有一个或者多个程序,当上述一个或者多个程序被该终端或服务器执行时,使得该服务器:依据录制的目标歌曲的音频内容和图像内容,生成目标歌曲的音视频;若检测到保存事件,则保存音视频,并根据预设的片段处理规则对音视频进行处理,生成精选片段;若检测到发布事件,则将精选片段添加到视频列表界面中进行展示。
可以以一种或多种程序设计语言或其组合来编写用于执行本公开的操作的计算机程序代码,上述程序设计语言包括面向对象的程序设计语言—诸如Java、Smalltalk、C++,还包括常规的过程式程序设计语言—诸如“C”语言或类似的程序设计语言。程序代码可以完全地在用户计算机上执行、部分地在用户计算机上执行、作为一个独立的软件包执行、部分在用户计算机上部分在远程计算机上执行、或者完全在远程计算机或服务器上执行。在涉及远程计算机的情形中,远程计算机可以通过任意种类的网络——包括局域网(Local Area Network,LAN)或广域网(Wide Area Network,WAN)—连接到用户计算机,或者,可以连接到外部计算机(例如利用因特网服务提供商来通过因特网连接)。
附图中的流程图和框图,图示了按照本公开多种实施例的系统、方法和计算机程序产品的可能实现的体系架构、功能和操作。在这点上,流程图或框图中的每个方框可以代表一个模块、程序段、或代码的一部分,该模块、程序段、或代码的一部分包含一个或多个用于实现规定的逻辑功能的可执行指令。在有些作为替换的实现中,方框中所标注的功能也可以以不同于附图中所标注的顺序发生。例如,两个接连地表示的方框实际上可以基本并行地执行,它们有时也可以按相反的顺序执行,这依所涉及的功能而定。框图和/或流程图中的每个方框、以及框图和/或流程图中的方框的组合,可以用执行规定的功能或操作的专用的基于硬件的系统来实现,或者可以用专用硬件与计算机指令的组合来实现。
描述于本公开实施例中所涉及到的单元可以通过软件的方式实现,也可以 通过硬件的方式来实现。其中,单元的名称在一种情况下并不构成对该单元本身的限定。
本文中以上描述的功能可以至少部分地由一个或多个硬件逻辑部件来执行。例如,非限制性地,可以使用的示范类型的硬件逻辑部件包括:现场可编程门阵列(Field Programmable Gate Array,FPGA)、专用集成电路(Application Specific Integrated Circuit,ASIC)、专用标准产品(Application Specific Standard Parts,ASSP)、片上系统(System-on-a-Chip,SOC)、复杂可编程逻辑设备(Complex Programmable Logic Device,CPLD)等等。
在本公开的上下文中,机器可读介质可以是有形的介质,机器可读介质可以包含或存储以供指令执行系统、装置或设备使用或与指令执行系统、装置或设备结合地使用的程序。机器可读介质可以是机器可读信号介质或机器可读储存介质。机器可读介质可以包括电子的、磁性的、光学的、电磁的、红外的、或半导体系统、装置或设备,或者上述内容的任何合适组合。机器可读存储介质包括基于一个或多个线的电气连接、便携式计算机盘、硬盘、RAM、ROM、EPROM或快闪存储器、光纤、CD-ROM、光学储存设备、磁储存设备、或上述内容的任何合适组合。

Claims (10)

  1. 一种精选片段处理方法,包括:
    依据录制的目标歌曲的音频内容和图像内容,生成所述目标歌曲的音视频;
    在检测到保存事件的情况下,保存所述音视频,并从所述音视频中确定一个预设时长的片段,作为精选片段;
    在检测到发布事件的情况下,将所述精选片段添加到视频列表界面中进行展示。
  2. 根据权利要求1所述的方法,其中,所述从所述音视频中确定一个预设时长的片段,作为精选片段,包括:
    将所述音视频中演唱第一句歌词的时间,或进入高潮片段的时间作为所述精选片段的起始时间;
    依据所述起始时间以及所述预设时长,确定所述精选片段的结束时间;
    将所述音视频中所述起始时间和所述结束时间之间的片段,作为所述精选片段。
  3. 根据权利要求1所述的方法,其中,所述从所述音视频中确定一个预设时长的片段,作为精选片段,包括:
    从所述音视频中选择一个第一预设时长的片段,作为第一片段;
    获取一个第二预设时长的片段,作为第二片段;其中,所述第一预设时长大于所述第二预设时长;
    对所述第一片段和所述第二片段进行拼接处理,生成所述精选片段。
  4. 根据权利要求3所述的方法,其中,所述第二片段为介绍片段,用于对所述第一片段的内容进行介绍。
  5. 根据权利要求1所述的方法,其中,所述从所述音视频中确定一个预设时长的片段,作为所述精选片段,包括:
    向用户展示精选片段制作界面;
    依据用户在所述精选片段制作界面上的操作,确定所述预设时长的起始时间和结束时间作为所述精选片段的起始时间和结束时间;
    将所述音视频中所述起始时间和所述结束时间之间的片段,作为所述精选片段。
  6. 根据权利要求1所述的方法,其中,所述精选片段中嵌入有进入所述音视频的接口,所述接口以物理按钮或悬浮球形式体现。
  7. 根据权利要求1所述的方法,其中,所述将所述精选片段添加到视频列表界面中进行展示,包括:
    将所述精选片段添加至所述视频列表界面,并确定包括所述精选片段的视频列表界面的加载信息;
    依据所述加载信息,加载包括所述精选片段的视频列表界面并展示所述精选片段。
  8. 一种精选片段处理装置,包括:
    音视频生成模块,设置为依据录制的目标歌曲的音频内容和图像内容,生成所述目标歌曲的音视频;
    精选片段生成模块,设置为在检测到保存事件的情况下,保存所述音视频,并从所述音视频中确定一个预设时长的片段,作为精选片段;
    添加展示模块,设置为在检测到发布事件的情况下,将所述精选片段添加到视频列表界面中进行展示。
  9. 一种电子设备,包括:
    一个或多个处理器;
    存储器,设置为存储一个或多个程序;
    所述一个或多个程序被所述一个或多个处理器执行,使得所述一个或多个处理器实现如权利要求1-7中任一项所述的精选片段处理方法。
  10. 一种可读介质,所述可读介质上存储有计算机程序,所述计算机程序被处理器执行时实现如权利要求1-7中任一项所述的精选片段处理方法。
PCT/CN2020/091071 2019-06-27 2020-05-19 精选片段处理方法、装置、电子设备及可读介质 Ceased WO2020259130A1 (zh)

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
CN201910569485.9 2019-06-27
CN201910569485.9A CN110312162A (zh) 2019-06-27 2019-06-27 精选片段处理方法、装置、电子设备及可读介质

Publications (1)

Publication Number Publication Date
WO2020259130A1 true WO2020259130A1 (zh) 2020-12-30

Family

ID=68077018

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/CN2020/091071 Ceased WO2020259130A1 (zh) 2019-06-27 2020-05-19 精选片段处理方法、装置、电子设备及可读介质

Country Status (2)

Country Link
CN (1) CN110312162A (zh)
WO (1) WO2020259130A1 (zh)

Families Citing this family (7)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
CN110312162A (zh) * 2019-06-27 2019-10-08 北京字节跳动网络技术有限公司 精选片段处理方法、装置、电子设备及可读介质
CN110225409A (zh) * 2019-07-19 2019-09-10 北京字节跳动网络技术有限公司 音视频播放方法、装置、电子设备及可读介质
CN111246244B (zh) * 2020-02-04 2023-05-23 北京贝思科技术有限公司 集群内快速分析处理音视频的方法、装置及电子设备
CN112349272A (zh) * 2020-10-15 2021-02-09 北京捷通华声科技股份有限公司 语音合成方法、装置、存储介质及电子装置
CN114661943A (zh) * 2022-05-21 2022-06-24 中科云策(深圳)科技成果转化信息技术有限公司 会议信息存储管理系统
CN115103219A (zh) * 2022-07-01 2022-09-23 抖音视界(北京)有限公司 音频发布方法、装置和计算机可读存储介质
CN116033096B (zh) * 2022-07-08 2023-10-20 荣耀终端有限公司 一种画面内容配音方法、装置及终端设备

Citations (9)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
WO2007109640A2 (en) * 2006-03-20 2007-09-27 Intension Inc. Methods of enhancing media content narrative
CN103338345A (zh) * 2013-06-09 2013-10-02 福建星网视易信息系统有限公司 演唱时拍摄图像或视频的方法与装置
CN104811787A (zh) * 2014-10-27 2015-07-29 深圳市腾讯计算机系统有限公司 游戏视频录制方法及装置
CN107910024A (zh) * 2017-10-10 2018-04-13 深圳市金立通信设备有限公司 一种录制数据的方法、终端及计算机可读存储介质
CN108833969A (zh) * 2018-06-28 2018-11-16 腾讯科技(深圳)有限公司 一种直播流的剪辑方法、装置以及设备
CN109040773A (zh) * 2018-07-10 2018-12-18 武汉斗鱼网络科技有限公司 一种视频改进方法、装置、设备及介质
CN109218746A (zh) * 2018-11-09 2019-01-15 北京达佳互联信息技术有限公司 获取视频片段的方法、装置和存储介质
CN109361954A (zh) * 2018-11-02 2019-02-19 腾讯科技(深圳)有限公司 视频资源的录制方法、装置、存储介质及电子装置
CN110312162A (zh) * 2019-06-27 2019-10-08 北京字节跳动网络技术有限公司 精选片段处理方法、装置、电子设备及可读介质

Family Cites Families (3)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
CN105872806A (zh) * 2016-05-05 2016-08-17 苏州花坞信息科技有限公司 一种在线视频播放方法
CN107995515B (zh) * 2017-11-30 2021-01-29 华为技术有限公司 信息提示的方法及装置
CN109361945A (zh) * 2018-10-18 2019-02-19 广州市保伦电子有限公司 一种快速传输及同步的会议视听系统及其控制方法

Patent Citations (9)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
WO2007109640A2 (en) * 2006-03-20 2007-09-27 Intension Inc. Methods of enhancing media content narrative
CN103338345A (zh) * 2013-06-09 2013-10-02 福建星网视易信息系统有限公司 演唱时拍摄图像或视频的方法与装置
CN104811787A (zh) * 2014-10-27 2015-07-29 深圳市腾讯计算机系统有限公司 游戏视频录制方法及装置
CN107910024A (zh) * 2017-10-10 2018-04-13 深圳市金立通信设备有限公司 一种录制数据的方法、终端及计算机可读存储介质
CN108833969A (zh) * 2018-06-28 2018-11-16 腾讯科技(深圳)有限公司 一种直播流的剪辑方法、装置以及设备
CN109040773A (zh) * 2018-07-10 2018-12-18 武汉斗鱼网络科技有限公司 一种视频改进方法、装置、设备及介质
CN109361954A (zh) * 2018-11-02 2019-02-19 腾讯科技(深圳)有限公司 视频资源的录制方法、装置、存储介质及电子装置
CN109218746A (zh) * 2018-11-09 2019-01-15 北京达佳互联信息技术有限公司 获取视频片段的方法、装置和存储介质
CN110312162A (zh) * 2019-06-27 2019-10-08 北京字节跳动网络技术有限公司 精选片段处理方法、装置、电子设备及可读介质

Also Published As

Publication number Publication date
CN110312162A (zh) 2019-10-08

Similar Documents

Publication Publication Date Title
US12483746B2 (en) Display method, apparatus, device and storage medium
CN112911379B (zh) 视频生成方法、装置、电子设备和存储介质
WO2020259130A1 (zh) 精选片段处理方法、装置、电子设备及可读介质
CN113365134B (zh) 音频分享方法、装置、设备及介质
WO2020253806A1 (zh) 展示视频的生成方法、装置、设备及存储介质
EP4124052B1 (en) Video production method and apparatus, and device and storage medium
CN115633206A (zh) 媒体内容展示方法、装置、设备及存储介质
WO2021012764A1 (zh) 音视频播放方法、装置、电子设备及可读介质
WO2020259133A1 (zh) 录制热门片段方法、装置、电子设备和可读介质
WO2023051293A1 (zh) 一种音频处理方法、装置、电子设备和存储介质
WO2020220776A1 (zh) 图片类评论数据的展示方法、装置、设备及介质
CN115981769A (zh) 页面显示方法、装置、设备、计算机可读存储介质及产品
WO2023088484A1 (zh) 用于多媒体资源剪辑场景的方法、装置、设备及存储介质
CN107450874A (zh) 一种多媒体数据双屏播放方法及系统
CN116366918A (zh) 媒体内容生成方法、装置、设备、可读存储介质及产品
WO2024032635A1 (zh) 媒体内容获取方法、装置、设备、可读存储介质及产品
WO2022218109A1 (zh) 交互方法, 装置, 电子设备及计算机可读存储介质
WO2020207101A1 (zh) 视频信息同步显示方法、装置、终端设备及计算机可读存储介质
WO2020224294A1 (zh) 用于处理信息的系统、方法和装置
CN117459776A (zh) 一种多媒体内容的播放方法、装置、电子设备及存储介质
CN110381356B (zh) 音视频生成方法、装置、电子设备及可读介质
WO2023174073A1 (zh) 视频生成方法、装置、设备、存储介质和程序产品
WO2022179522A1 (zh) 推荐视频展示方法、装置、介质及电子设备
WO2022257777A1 (zh) 多媒体处理方法、装置、设备及介质
CN114677738A (zh) Mv录制方法、装置、电子设备及计算机可读存储介质

Legal Events

Date Code Title Description
121 Ep: the epo has been informed by wipo that ep was designated in this application

Ref document number: 20831917

Country of ref document: EP

Kind code of ref document: A1

NENP Non-entry into the national phase

Ref country code: DE

32PN Ep: public notification in the ep bulletin as address of the adressee cannot be established

Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205A DATED 21/04/2022)

122 Ep: pct application non-entry in european phase

Ref document number: 20831917

Country of ref document: EP

Kind code of ref document: A1