WO2022017203A1 - 听写交互方法、装置和电子设备 - Google Patents
听写交互方法、装置和电子设备 Download PDFInfo
- Publication number
- WO2022017203A1 WO2022017203A1 PCT/CN2021/105611 CN2021105611W WO2022017203A1 WO 2022017203 A1 WO2022017203 A1 WO 2022017203A1 CN 2021105611 W CN2021105611 W CN 2021105611W WO 2022017203 A1 WO2022017203 A1 WO 2022017203A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- dictation
- user
- target
- terminal
- configuration interface
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Images
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F9/00—Arrangements for program control, e.g. control units
- G06F9/06—Arrangements for program control, e.g. control units using stored programs, i.e. using an internal store of processing equipment to receive or retain programs
- G06F9/44—Arrangements for executing specific programs
- G06F9/451—Execution arrangements for user interfaces
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F9/00—Arrangements for program control, e.g. control units
- G06F9/06—Arrangements for program control, e.g. control units using stored programs, i.e. using an internal store of processing equipment to receive or retain programs
- G06F9/44—Arrangements for executing specific programs
- G06F9/445—Program loading or initiating
- G06F9/44505—Configuring for program initiating, e.g. using registry, configuration files
- G06F9/4451—User profiles; Roaming
-
- G—PHYSICS
- G09—EDUCATION; CRYPTOGRAPHY; DISPLAY; ADVERTISING; SEALS
- G09B—EDUCATIONAL OR DEMONSTRATION APPLIANCES; APPLIANCES FOR TEACHING, OR COMMUNICATING WITH, THE BLIND, DEAF OR MUTE; MODELS; PLANETARIA; GLOBES; MAPS; DIAGRAMS
- G09B5/00—Electrically-operated educational appliances
- G09B5/08—Electrically-operated educational appliances providing for individual presentation of information to a plurality of student stations
- G09B5/14—Electrically-operated educational appliances providing for individual presentation of information to a plurality of student stations with provision for individual teacher-student communication
Definitions
- the present disclosure relates to the field of Internet technologies, and in particular, to a dictation interaction method, apparatus, and electronic device.
- dictation is an important learning method. Traditional dictation requires teachers and students to be together, the teacher reads aloud, and the students write what they hear. Traditional dictation has obvious limitations in space or time.
- an embodiment of the present disclosure provides a dictation interaction method, which is applied to a first terminal.
- the method includes: displaying a configuration interface, and determining target dictation materials and dictation participants through the configuration interface, wherein the The configuration interface is used for the dictation initiating user to configure the target dictation material and the dictation participant user identifier; the voice file corresponding to the target dictation material is sent to the second terminal corresponding to the dictation participant user identifier, wherein the dictation participant user is based on the voice file. file for dictation.
- an embodiment of the present disclosure provides a dictation interaction device, applied to a first terminal, the device includes: a display unit, configured to display a configuration interface, and determine a target dictation material and a dictation participant user through the configuration interface , wherein the configuration interface is used for the dictation initiating user to configure the target dictation material and the dictation participant user identifier; the sending unit is used to send the voice file corresponding to the target dictation material to the second terminal corresponding to the dictation participant user identifier , wherein the dictation participating user performs dictation based on the voice file.
- embodiments of the present disclosure provide an electronic device, including: one or more processors; and a storage device for storing one or more programs, when the one or more programs are stored by the one or more programs The one or more processors execute, so that the one or more processors implement the dictation interaction method as described in the first aspect.
- an embodiment of the present disclosure provides a computer-readable medium on which a computer program is stored, and when the program is executed by a processor, implements the steps of the dictation interaction method described in the first aspect.
- the dictation initiating user configures the target dictation material and the dictation participant user identifier; then, the voice file corresponding to the target dictation material is sent to the dictation participant user.
- the dictation-initiating user can arrange dictation tasks online, and the dictation process is not limited by space.
- FIG. 1 is a flowchart of one embodiment of a dictation interaction method according to the present disclosure
- FIG. 2 is a schematic diagram of an exemplary configuration interface according to the present disclosure
- FIG. 3 is an exemplary schematic diagram of a second terminal displaying notification information according to the present disclosure
- FIG. 4 is an exemplary schematic diagram of a first terminal showing dictation progress information according to the present disclosure
- FIG. 5 is an exemplary schematic diagram of the first terminal of the present disclosure correcting the dictation image information
- FIG. 6 is an exemplary schematic diagram of a second terminal showing a correction result according to the present disclosure
- FIG. 7 is an exemplary schematic diagram showing interpretation information of a dictation item according to the present disclosure.
- FIG. 8 is a schematic structural diagram of an embodiment of a dictation interaction device according to the present disclosure.
- FIG. 9 is an exemplary system architecture to which the dictation interaction method of an embodiment of the present disclosure may be applied.
- FIG. 10 is a schematic diagram of a basic structure of an electronic device provided according to an embodiment of the present disclosure.
- the term “including” and variations thereof are open-ended inclusions, ie, "including but not limited to”.
- the term “based on” is “based at least in part on.”
- the term “one embodiment” means “at least one embodiment”; the term “another embodiment” means “at least one additional embodiment”; the term “some embodiments” means “at least some embodiments”. Relevant definitions of other terms will be given in the description below.
- FIG. 1 shows a flow of an embodiment of a dictation interaction method according to the present disclosure.
- the dictation interaction method includes the following steps:
- step 101 a configuration interface is displayed, and a target dictation material and a dictation participant user are determined through the configuration interface.
- the execution body (for example, the first terminal) of the dictation interaction method may display a configuration interface.
- the above-mentioned configuration interface may be used for the dictation initiating user to configure the target dictation material and the dictation participant user identification.
- the dictation initiating user may be the user who initiates the dictation, in other words, the user who arranges the dictation task.
- the target dictation material may be the content to be dictated.
- the specific form of the target dictation material is not limited.
- the target dictation material may be characters, words, sentences, paragraphs, and the like.
- the dictation participation user identification may indicate the user listening to the target dictation material for the writing operation.
- the number of user IDs for dictation participation may be one, or at least two, which are not limited herein.
- a dictation initiating user may be a teacher and a dictation participating user may be a student.
- the above configuration interface may be one interface or multiple interfaces.
- configuring the target dictation material and configuring the dictation participant user identification can be performed in the same interface or in different interfaces. If performed in different interfaces, these different interfaces can be collectively referred to as configuration interfaces.
- Step 102 Send the voice file corresponding to the target dictation material to the terminal device corresponding to the dictation participating user ID.
- the above-mentioned execution body may send the voice file corresponding to the above-mentioned target dictation material to the terminal device corresponding to the dictation participating user identification.
- the dictation participating user indicated by the dictation participating user identifier may perform dictation based on the above-mentioned voice file.
- the above configuration interface may set a dictation initiation confirmation control, and the above dictation initiation user may trigger the dictation initiation confirmation control after configuring the target dictation material and the dictation participant user ID. Then, based on the configuration of the dictation initiating user, the above-mentioned execution body may send the voice file corresponding to the target dictation material to the terminal device corresponding to the dictation participant user.
- the above-mentioned voice file may be sent directly by the above-mentioned executive body, or may be sent by the above-mentioned executive body indirectly.
- the above-mentioned execution body may send the target dictation material to the server, and the server may obtain the generated voice file or generate the voice file according to the target dictation material.
- the terminal device (which may be referred to as the second terminal device in this application) corresponding to the user identification participating in the dictation can play the voice file. Dictation participating users listen to the voice played by the terminal device and perform dictation.
- the above-mentioned second terminal device may start playing the voice file upon receiving it, or the user participating in the dictation may select the playing time before playing.
- FIG. 2 shows an exemplary schematic diagram of a configuration interface.
- the dictation initiating user configures the target dictation material and the dictation participant user identifier; then, the voice file corresponding to the target dictation material is sent to the dictation participant user.
- the dictation-initiating user can arrange dictation tasks online, and the dictation process is not limited by space.
- the above-mentioned configuration interface may include at least one of the following, but is not limited to: candidate dictation content identification and dictation content supplementary area.
- the candidate dictation content identification may indicate pre-stored candidate dictation content.
- the above-mentioned dictation initiating user may determine the target dictation material by selecting candidate dictation content identifiers.
- the execution body may determine the candidate dictation content indicated by the candidate dictation identifier targeted by the first selection operation as the target dictation material.
- the above-mentioned dictation initiating user can configure the target dictation material by performing a series of selection operations.
- the user can be initiated by the above-mentioned dictation, and can select subjects, teaching materials, and texts in sequence.
- the dictation initiating user selects the text ID
- the pre-associated vocabulary of the text ID can be used as the candidate dictation content for the user to select the target dictation material from the candidate dictation content.
- the three dictation item identifiers of "green”, “silk” and “stripe” are understood as candidate dictation content identifiers.
- the user who initiates the dictation can select the target dictation material from the three dictation items of "green”, “silk” and "stripe”.
- the configuration interface described above includes a dictation content supplement area.
- the above step 101 may include: in response to acquiring the dictation supplementary content input in the dictation content supplementary area, determining the dictation supplementary content as the target dictation material.
- the above-mentioned dictation initiating user can input desired dictation material in the dictation content supplement area, and then trigger the confirmation control to confirm that the input dictation material is determined as the material finally input into the above-mentioned dictation content supplement area.
- the above executive body may also use the dictation material input in the dictation content supplement area as the target dictation material.
- the dictation initiating user can flexibly arrange the dictation task.
- the box under “Select dictation content” in FIG. 2 can be understood as a supplementary area of dictation content.
- the "scissors" in the dictation supplement area can be dictation supplements.
- the dictation mode information may indicate the dictation mode.
- the dictation mode information may include, but is not limited to, at least one of the following: the number of times of repeated word reading, the time interval between reading words, and dictation order indication information.
- the number of repeated readings may indicate the number of repeated playbacks of a single dictation item.
- the reading interval duration may indicate the duration between the end of playing the word and the start of playing the next dictation item of the dictation item.
- the dictation order indication information may indicate the playback order of the dictation items in the dictation material.
- the playback order may be random or sequential.
- the same out-of-order mode can be used for all dictation participating users.
- the order of the dictation items in the multiple dictation content images received by the dictation initiating user is consistent, so that the correction efficiency of the dictation initiating user can be improved.
- the above method may include: acquiring, through the configuration interface, the dictation mode information configured by the dictation initiating user for the target dictation material.
- the terminal device may use the dictation mode indicated by the dictation mode information to play the voice file.
- dictation methods may change the difficulty of the dictation task.
- the greater the number of repeated word readings the lower the dictation difficulty; the longer the word reading interval, the lower the dictation difficulty; the sequential dictation is lower than the disordered dictation, and the dictation difficulty is lower.
- the dictation initiating user can flexibly set the dictation mode information conforming to the actual application for this dictation task, that is, flexibly set the difficulty of the dictation task.
- FIG. 2 shows a schematic diagram of a scenario for configuring dictation mode information.
- Word reading interval, dictation order, and word reading times can be configured as dictation mode information.
- the target dictation material includes at least one dictation item.
- Dictation items can be words, words, sentences, or paragraphs.
- the voice file corresponding to the target dictation material can be determined in the following manner: for each dictation item in the target dictation material, it is determined whether the voice file corresponding to the dictation item has been generated; in response to determining that it has not been generated, the voice file corresponding to the dictation item is synthesized. Then, the synthesized speech file can be stored in the speech file corresponding to the target dictation material.
- the above step of determining the voice file corresponding to the target dictation material may be executed by the above-mentioned execution body, or may be executed by the above-mentioned server supporting the above-mentioned execution body.
- each dictation item has a corresponding language file, and if not, synthesizing a voice file, the workload of the dictation initiating user can be reduced, the steps of arranging dictation tasks can be reduced, and the efficiency of dictation task arrangement can be improved.
- the above-mentioned second terminal (the terminal device corresponding to the dictation participant user identification) can receive the dictation notification, and can display the dictation notification.
- the above-mentioned dictation notification is used to notify the dictation participating user of the corresponding dictation task.
- FIG. 3 shows a dictation notification displayed by the terminal.
- the dictation notice can include "The teacher has assigned you a dictation task” and “The texts such as “Two Ancient Poems” need to be dictated, hurry up and complete it”; and, the dictation notice can include confirmation controls (in Figure 3 marked "View Now”).
- the second terminal may present a dictation notification and initiate a dictation process in response to detecting a predefined dictation start operation.
- the specific content of the predefined dictation start operation can be set according to the actual application scenario, which is not limited here.
- the dictation start timing of the second terminal may be the time when the dictation notification is received, the dictation start time configured by the dictation initiating user, or the time selected by the dictation participant user.
- a predefined dictation start operation may include a trigger operation for a dictation notification.
- trigger a dictation notification, and the dictation process begins.
- a dictation start confirmation control may be displayed.
- the dictation participant user triggers the dictation start confirmation control, which can be used as a predefined dictation start operation.
- an application can set a dictation task viewing entry, and users can view their own dictation tasks from the dictation task viewing entry.
- Unfinished dictation tasks can be associated with a dictation start confirmation control.
- the trigger operation of the dictation participating user on the dictation start confirmation control can be used as a predefined dictation start operation. In other words, the dictation participating user triggers the unfinished dictation task, and the dictation task can be started.
- the above-mentioned second execution body may play a voice file, and capture a user image of a dictation participant user.
- the user image of the dictation participant user may include an image of any body part of the dictation participant user during the dictation process.
- the upper body image of the dictation participant user may be collected, and the hand image of the dictation participant user writing may also be collected.
- the above-mentioned dictation participant user may write on the second terminal, or may use pen and paper to write.
- the above-mentioned first execution body may include a camera, and the camera may directly aim at the handwriting position of the dictation participant user for image capture, and may also perform image capture on the handwriting position of the dictation participant user by setting the optical path (eg, setting a reflector). That is, the second terminal may receive the writing image of the dictation participant user through the camera.
- the above-mentioned second terminal may also receive the writing image of the dictation participant user through the display screen.
- the second terminal may acquire the writing process of the dictation participant user by recording the screen.
- receiving the writing content input by the user on the display screen through the second terminal can avoid the step of the user preparing pen and paper for dictation, and reduce the restriction of the dictation tool on the implementation of dictation. Therefore, users participating in dictation can start dictation anytime and anywhere, and the efficiency of dictation can be improved.
- the above-mentioned second terminal may transmit the user image of the dictation participant user to a preset electronic device (for example, a server) in real time. Moreover, the second terminal can feed back the dictation progress information to the first terminal in real time. If the first terminal detects the operation of acquiring the user image, the first terminal can pull the stream from the above-mentioned preset electronic device to display the user image (or video).
- a preset electronic device for example, a server
- the second terminal can feed back the dictation progress information to the first terminal in real time. If the first terminal detects the operation of acquiring the user image, the first terminal can pull the stream from the above-mentioned preset electronic device to display the user image (or video).
- the above-mentioned second terminal may capture the dictation content image in response to determining that the dictation ends.
- the above-mentioned second terminal may determine that the dictation ends in response to detecting a trigger operation for the preset dictation end control.
- the above-mentioned second terminal may determine that the dictation ends in response to a preset dictation duration elapsed from the start of the dictation.
- the preset dictation duration may be preset according to the target dictation material.
- the above-mentioned second terminal may determine that the dictation ends in response to a preset time period elapsed after the completion of the playback of the target dictation material at seven o'clock. As an example, when the target dictation material is played for 5 seconds, it can be determined that the dictation ends.
- the above-mentioned second terminal may suspend the image acquisition function in response to the end of the dictation.
- the above-mentioned second terminal may display guiding information, such as "please put the dictation content on the screen", so as to guide the dictation participant user to take an image of the dictation content.
- the above-mentioned second terminal may acquire an image of the written dictation content of the dictation participant user.
- the dictation content image collected by the second terminal is generally clearer than the image collected during the dictation process.
- the dictation content image collected at the end of the dictation is used as the basis for the correction of the user who initiates the dictation, which can avoid the reduction of correction efficiency and the inaccuracy of correction caused by unclear images.
- each dictation task related to the class can be displayed in a class.
- the three-year Chinese dictation task and the English dictation task can be shown.
- each dictation task initiated by the dictation initiating user may be displayed in units of the dictation initiating user.
- Mr. Li it is possible to display the Chinese dictation task initiated by Mr. Li for the first class of the third year, and the Chinese dictation task initiated by Mr. Li for the second class of the third year.
- overall execution progress information for the dictation task may be displayed.
- the overall execution progress information may include at least one of the following, but is not limited to: the total number of dictation participating users of the dictation task, the number of dictation participating users who have completed the dictation task, and the number of dictation participating users who have not completed the dictation task.
- the above-mentioned execution body may display the dictation progress information of each dictation participant user.
- FIG. 4 shows an application scenario in which the execution subject displays the dictation progress information for each dictation participation.
- the dictation progress information "dictation completed” can be displayed correspondingly; for the dictation participant user “Li Si”, the dictation progress information “dictation 20%” can be displayed correspondingly; for the dictation participant user “Dictation 20%”
- the dictation progress information "Completed” can be displayed correspondingly; for the dictation participant “Song Liu”, the dictation progress information "Not Started” can be displayed correspondingly.
- the dictation initiating user can obtain the execution status of the assigned dictation task in time, and thus, the dictation initiating user can use this as a basis to timely participate in the dictation that has not started. Users are reminded to improve the efficiency of dictation interaction.
- the above-mentioned execution body may acquire and play the images during the dictation process for the dictation participant users whose dictation is in progress or completed.
- the dictation initiating user can view the user image or video of the dictation process.
- the user image of the dictation participant user in the dictation process can be recorded. Therefore, the dictation initiating user can view the user image in the dictation process, so that the dictation initiating user can supervise the dictation process in real time or non-real time, and improve the dictation efficiency.
- the method includes: displaying the user image and the writing image having the associated relationship.
- the writing image may be handwritten by the user through a pen and paper, and collected by the second terminal; optionally, the writing image may also be handwritten by the user through the display screen of the second terminal, and the second terminal may acquire by recording the screen or the like. .
- the actual dictation process can be shown to the dictation initiating user by showing the associated user image and the writing image. Therefore, the dictation initiating user can obtain more user information in the actual dictation process through the writing image and the user image in the actual dictation process, so as to effectively supervise the dictation process.
- the above-mentioned execution body may display the dictation content image and the target dictation material; and generate a correction result of the dictation content image according to the second selection operation on the target dictation material.
- displaying the dictation content image and the dictation material side by side can facilitate the selection of the wrong part of the dictation by the user who corrects the assignment (usually the user who initiates the dictation).
- FIG. 5 shows a schematic diagram of a correction process.
- Zhang San's dictation content image is shown, and the target dictation material is shown.
- Zhang San's homework he mistakenly spelled " ⁇ " as "stripe", so that the correcting user can select " ⁇ " (indicated by shading) in the target dictation material, and the result of homework correction can be obtained.
- the method further includes returning a correction result to the target user to be corrected, wherein the correction result includes the target dictation material and error item indication information.
- the error item indication information may indicate an error item in the dictation content image.
- the second terminal in response to detecting the trigger operation for the dictation item, acquires interpretation information of the dictation item targeted by the trigger operation, and displays the acquired interpretation information.
- interpretation information can be used to interpret dictation items.
- Interpretation information can be stored in advance.
- the source of interpretation information may include, but is not limited to, at least one of the following: a dictionary, a textbook, and the like.
- FIG. 6 shows a schematic diagram of a scenario where the second terminal displays the correction result.
- the answer of Zhang San that is, the logged-in user of the second terminal
- my answer can be displayed.
- the correction result is displayed, and the correction result can take various forms.
- the correction result can be displayed in the form of adding error item indication information to the dictation item of the target dictation material.
- FIG. 7 shows a schematic diagram of a scenario where the second terminal displays interpretation information. Zhang San can click on the "sash” shown in shadow in FIG. 6 , and then the second terminal can display the explanation information of the “sash” shown in FIG. 6 .
- associating the dictation items in the target dictation materials with the explanation information can enable the dictation participants to quickly obtain the detailed information of the dictation items, so as to effectively consolidate the learning of the dictation participants in a timely manner. Effect.
- the present disclosure provides an embodiment of a dictation interaction device, the device embodiment corresponds to the method embodiment shown in FIG. 1 , and the device may specifically be Used in various electronic devices.
- the dictation interaction apparatus in this embodiment includes: a presentation unit 801 and a sending unit.
- the display unit is used to display the configuration interface, and determine the target dictation material and the dictation participant user through the configuration interface, wherein the configuration interface is used for the dictation initiating user to configure the target dictation material and the dictation participant user identifier; the sending unit , which is used to send the voice file corresponding to the target dictation material to the second terminal corresponding to the dictation participant user ID, where the dictation participant user performs dictation based on the voice file.
- step 101 and step 102 in the corresponding embodiment of FIG. 8 , which are not repeated here. Repeat.
- the configuration interface includes at least one of the following: candidate dictation content identification and dictation content supplemental areas; and the presentation configuration interface, and determining target dictation materials and dictation participating users through the configuration interface, including: Displaying the candidate dictation content identifier, and in response to detecting the first selection operation for the candidate dictation content identifier, determining the candidate dictation content indicated by the candidate dictation identifier targeted by the first selection operation as the target dictation material; displaying the dictation content supplementary area, and in response to acquiring the dictation supplementary content entered in the dictation content supplementary area, determining the dictation supplementary content as the target dictation material.
- the configuration interface includes a dictation mode information configuration area; and the apparatus is further configured to: obtain, through the configuration interface, the dictation mode information configured by the dictation initiating user for the target dictation material, wherein the The terminal device uses the dictation mode indicated by the dictation mode information to play the voice file, and the dictation mode information includes at least one of the following: the number of repeated word readings, the time interval between reading words, and the dictation order indication information.
- the target dictation material includes at least one dictation item
- the voice file corresponding to the target dictation material is determined by the following method: for each dictation item in the target dictation material, it is determined whether a corresponding dictation item has been generated A voice file; in response to determining that it is not generated, synthesize a voice file corresponding to the dictation item.
- the second terminal presents a dictation notification.
- the second terminal starts an operation in response to a predefined dictation, plays the voice file, and captures a user image of a dictation participant user.
- the second terminal receives the writing image of the dictation participant user through a display screen and/or a camera.
- the second terminal captures the dictation content image in response to determining that the dictation is over.
- the apparatus is further configured to: display the dictation progress information of each dictation participant user.
- the apparatus is further configured to: display a user image during the dictation process for the dictation participant user who is in the process of dictation or who has completed the dictation.
- the apparatus is further configured to: display the user image and the writing image with the associated relationship.
- the device is further configured to: display the dictation content image and the target dictation material; generate a correction result of the dictation content image according to the second selection operation on the dictation item in the target dictation material, wherein the correction result includes the target Dictate material and error item instructions.
- the second terminal displays the correction result, and in response to detecting the trigger operation for the dictation item, acquires and displays explanation information of the dictation item targeted by the trigger operation.
- FIG. 9 illustrates an exemplary system architecture to which the dictation interaction method according to an embodiment of the present disclosure may be applied.
- the system architecture may include terminal devices 901 , 902 , and 903 , a network 904 , and a server 905 .
- the network 904 is a medium used to provide a communication link between the terminal devices 901 , 902 , 903 and the server 905 .
- Network 904 may include various connection types, such as wired, wireless communication links, or fiber optic cables, among others.
- the terminal devices 901, 902, 903 can interact with the server 905 through the network 904 to receive or send messages and the like.
- Various client applications may be installed on the terminal devices 901 , 902 and 903 , such as web browser applications, search applications, and news information applications.
- the client applications in the terminal devices 901, 902, and 903 can receive the user's instruction, and perform corresponding functions according to the user's instruction, for example, adding corresponding information to the information according to the user's instruction.
- the terminal devices 901, 902, and 903 may be hardware or software.
- the terminal devices 901, 902, and 903 can be various electronic devices that have a display screen and support web browsing, including but not limited to smart phones, tablet computers, e-book readers, MP3 players (Moving Picture Experts Group Audio Layer III, Moving Picture Experts Compression Standard Audio Layer 3), MP4 (Moving Picture Experts Group Audio Layer IV, Moving Picture Experts Compression Standard Audio Layer 4) Players, Laptops and Desktops, etc.
- the terminal devices 901, 902, and 903 are software, they can be installed in the electronic devices listed above. It can be implemented as a plurality of software or software modules (eg, software or software modules for providing distributed services), or can be implemented as a single software or software module. There is no specific limitation here.
- the server 905 may be a server that provides various services, such as receiving information acquisition requests sent by the terminal devices 901, 902, and 903, and acquiring display information corresponding to the information acquisition requests in various ways according to the information acquisition requests. And the relevant data of the displayed information is sent to the terminal devices 901 , 902 and 903 .
- the dictation interaction method provided by the embodiments of the present disclosure may be executed by a terminal device, and correspondingly, the dictation interaction apparatus may be set in the terminal devices 901 , 902 , and 903 .
- the dictation interaction method provided by the embodiment of the present disclosure may also be jointly executed by the terminal device and the server 905 , and accordingly, the dictation interaction apparatus may be provided in the terminal device and the server 905 .
- terminal devices, networks and servers in FIG. 9 are only illustrative. There can be any number of terminal devices, networks and servers according to implementation needs.
- FIG. 10 it shows a schematic structural diagram of an electronic device (eg, a terminal device or a server in FIG. 9 ) suitable for implementing an embodiment of the present disclosure.
- Terminal devices in the embodiments of the present disclosure may include, but are not limited to, such as mobile phones, notebook computers, digital broadcast receivers, PDAs (personal digital assistants), PADs (tablets), PMPs (portable multimedia players), vehicle-mounted terminals (eg, mobile terminals such as in-vehicle navigation terminals), etc., and stationary terminals such as digital TVs, desktop computers, and the like.
- the electronic device shown in FIG. 10 is only an example, and should not impose any limitation on the function and scope of use of the embodiments of the present disclosure.
- an electronic device may include a processing device (eg, a central processing unit, a graphics processor, etc.) 1001, which may be loaded into a random access memory according to a program stored in a read only memory (ROM) 1002 or from a storage device 1008
- the program in (RAM) 1003 executes various appropriate operations and processes.
- various programs and data required for the operation of the electronic device 1000 are also stored.
- the processing device 1001, the ROM 1002, and the RAM 1003 are connected to each other through a bus 1004.
- An input/output (I/O) interface 1005 is also connected to the bus 1004 .
- I/O interface 1005 input devices 1006 including, for example, a touch screen, touchpad, keyboard, mouse, camera, microphone, accelerometer, gyroscope, etc.; including, for example, a liquid crystal display (LCD), speakers, vibration
- An output device 1007 such as a computer
- a storage device 1008 including, for example, a magnetic tape, a hard disk, etc.
- Communication means 1009 may allow electronic devices to communicate wirelessly or by wire with other devices to exchange data. While FIG. 10 illustrates an electronic device having various means, it should be understood that not all of the illustrated means are required to be implemented or available. More or fewer devices may alternatively be implemented or provided.
- embodiments of the present disclosure include a computer program product comprising a computer program carried on a non-transitory computer readable medium, the computer program containing program code for performing the method illustrated in the flowchart.
- the computer program may be downloaded and installed from the network via the communication device 1009, or from the storage device 1008, or from the ROM 1002.
- the processing apparatus 1001 the above-mentioned functions defined in the methods of the embodiments of the present disclosure are executed.
- the computer-readable medium mentioned above in the present disclosure may be a computer-readable signal medium or a computer-readable storage medium, or any combination of the above two.
- the computer-readable storage medium can be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus or device, or a combination of any of the above. More specific examples of computer readable storage media may include, but are not limited to, electrical connections with one or more wires, portable computer disks, hard disks, random access memory (RAM), read only memory (ROM), erasable Programmable read only memory (EPROM or flash memory), fiber optics, portable compact disk read only memory (CD-ROM), optical storage devices, magnetic storage devices, or any suitable combination of the foregoing.
- a computer-readable storage medium may be any tangible medium that contains or stores a program that can be used by or in conjunction with an instruction execution system, apparatus, or device.
- a computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave with computer-readable program code embodied thereon. Such propagated data signals may take a variety of forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the foregoing.
- a computer-readable signal medium can also be any computer-readable medium other than a computer-readable storage medium that can transmit, propagate, or transport the program for use by or in connection with the instruction execution system, apparatus, or device .
- Program code embodied on a computer readable medium may be transmitted using any suitable medium including, but not limited to, electrical wire, optical fiber cable, RF (radio frequency), etc., or any suitable combination of the foregoing.
- the client and server can use any currently known or future developed network protocol such as HTTP (HyperText Transfer Protocol) to communicate, and can communicate with digital data in any form or medium Communication (eg, a communication network) interconnects.
- HTTP HyperText Transfer Protocol
- Examples of communication networks include local area networks (“LAN”), wide area networks (“WAN”), the Internet (eg, the Internet), and peer-to-peer networks (eg, ad hoc peer-to-peer networks), as well as any currently known or future development network of.
- the above-mentioned computer-readable medium may be included in the above-mentioned electronic device; or may exist alone without being assembled into the electronic device.
- the above-mentioned computer-readable medium carries one or more programs, and when the above-mentioned one or more programs are executed by the electronic device, the electronic device: displays a configuration interface, and determines a target dictation material and a dictation participant user through the configuration interface , wherein the configuration interface is used for the dictation initiating user to configure the target dictation material and the dictation participant user ID; the voice file corresponding to the target dictation material is sent to the second terminal corresponding to the dictation participant user ID, wherein the dictation participant ID The user dictates based on the voice file.
- Computer program code for performing operations of the present disclosure may be written in one or more programming languages, including but not limited to object-oriented programming languages—such as Java, Smalltalk, C++, and This includes conventional procedural programming languages - such as the "C" language or similar programming languages.
- the program code may execute entirely on the user's computer, partly on the user's computer, as a stand-alone software package, partly on the user's computer and partly on a remote computer, or entirely on the remote computer or server.
- the remote computer may be connected to the user's computer through any kind of network, including a local area network (LAN) or a wide area network (WAN), or may be connected to an external computer (eg, using an Internet service provider through Internet connection).
- LAN local area network
- WAN wide area network
- each block in the flowchart or block diagrams may represent a module, segment, or portion of code that contains one or more logical functions for implementing the specified functions executable instructions.
- the functions noted in the blocks may occur out of the order noted in the figures. For example, two blocks shown in succession may, in fact, be executed substantially concurrently, or the blocks may sometimes be executed in the reverse order, depending upon the functionality involved.
- each block of the block diagrams and/or flowchart illustrations, and combinations of blocks in the block diagrams and/or flowchart illustrations can be implemented in dedicated hardware-based systems that perform the specified functions or operations , or can be implemented in a combination of dedicated hardware and computer instructions.
- the units involved in the embodiments of the present disclosure may be implemented in a software manner, and may also be implemented in a hardware manner.
- the name of the unit does not constitute a limitation of the unit itself in some cases, for example, the display unit may also be described as a "unit for displaying a configuration interface".
- exemplary types of hardware logic components include: Field Programmable Gate Arrays (FPGAs), Application Specific Integrated Circuits (ASICs), Application Specific Standard Products (ASSPs), Systems on Chips (SOCs), Complex Programmable Logical Devices (CPLDs) and more.
- FPGAs Field Programmable Gate Arrays
- ASICs Application Specific Integrated Circuits
- ASSPs Application Specific Standard Products
- SOCs Systems on Chips
- CPLDs Complex Programmable Logical Devices
- a machine-readable medium may be a tangible medium that may contain or store a program for use by or in connection with the instruction execution system, apparatus or device.
- the machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium.
- Machine-readable media may include, but are not limited to, electronic, magnetic, optical, electromagnetic, infrared, or semiconductor systems, devices, or devices, or any suitable combination of the foregoing.
- machine-readable storage media would include one or more wire-based electrical connections, portable computer disks, hard disks, random access memory (RAM), read only memory (ROM), erasable programmable read only memory (EPROM or flash memory), fiber optics, compact disk read only memory (CD-ROM), optical storage, magnetic storage, or any suitable combination of the foregoing.
- RAM random access memory
- ROM read only memory
- EPROM or flash memory erasable programmable read only memory
- CD-ROM compact disk read only memory
- magnetic storage or any suitable combination of the foregoing.
Landscapes
- Engineering & Computer Science (AREA)
- Software Systems (AREA)
- Theoretical Computer Science (AREA)
- Physics & Mathematics (AREA)
- General Physics & Mathematics (AREA)
- General Engineering & Computer Science (AREA)
- Human Computer Interaction (AREA)
- Business, Economics & Management (AREA)
- Educational Administration (AREA)
- Educational Technology (AREA)
- User Interface Of Digital Computer (AREA)
Abstract
提供了一种听写交互方法、装置和电子设备。该方法包括:展示配置界面,以及通过配置界面确定目标听写材料和听写参与用户(101),其中,配置界面用于供听写发起用户配置目标听写材料和听写参与用户标识;将目标听写材料对应的语音文件,发送至听写参与用户标识对应的终端设备(102),其中,听写参与用户基于语音文件进行听写。由此,听写发起用户可以在线上布置听写任务,听写过程不受空间的限制。
Description
相关申请的交叉引用
本申请要求于2020年07月21日提交的,申请号为202010714796.2、发明名称为“听写交互方法、装置和电子设备”的中国专利申请的优先权,该申请的全文通过引用结合在本申请中。
本公开涉及互联网技术领域,尤其涉及一种听写交互方法、装置和电子设备。
在学习过程,尤其是语言学习过程中,听写是一种重要的学习方式。传统的听写,需要教师和学生在一起,由教师进行朗读,学生写出听到的内容。传统的听写在空间或时间上的限制非常明显。
发明内容
提供该公开内容部分以便以简要的形式介绍构思,这些构思将在后面的具体实施方式部分被详细描述。该公开内容部分并不旨在标识要求保护的技术方案的关键特征或必要特征,也不旨在用于限制所要求的保护的技术方案的范围。
第一方面,本公开实施例提供了一种听写交互方法,应用于第一终端,所述方法包括:展示配置界面,以及通过所述配置界 面确定目标听写材料和听写参与用户,其中,所述配置界面用于供听写发起用户配置目标听写材料和听写参与用户标识;将所述目标听写材料对应的语音文件,发送至听写参与用户标识对应的第二终端,其中,听写参与用户基于所述语音文件进行听写。
第二方面,本公开实施例提供了一种听写交互装置,应用于第一终端,所述装置包括:展示单元,用于展示配置界面,以及通过所述配置界面确定目标听写材料和听写参与用户,其中,所述配置界面用于供听写发起用户配置目标听写材料和听写参与用户标识;发送单元,用于将所述目标听写材料对应的语音文件,发送至听写参与用户标识对应的第二终端,其中,听写参与用户基于所述语音文件进行听写。
第三方面,本公开实施例提供了一种电子设备,包括:一个或多个处理器;存储装置,用于存储一个或多个程序,当所述一个或多个程序被所述一个或多个处理器执行,使得所述一个或多个处理器实现如第一方面所述的听写交互方法。
第四方面,本公开实施例提供了一种计算机可读介质,其上存储有计算机程序,该程序被处理器执行时实现如第一方面所述的听写交互方法的步骤。
本公开实施例提供的听写交互方法、装置和电子设备,通过先展示配置界面,供听写发起用户配置目标听写材料和听写参与用户标识;然后,将目标听写材料对应的语音文件,发送至听写参与用户。由此,听写发起用户可以在线上布置听写任务,听写过程不受空间的限制。
结合附图并参考以下具体实施方式,本公开各实施例的上述和其他特征、优点及方面将变得更加明显。贯穿附图中,相同或相似的附图标记表示相同或相似的元素。应当理解附图是示意性的,原件和元素不一定按照比例绘制。
图1是根据本公开的听写交互方法的一个实施例的流程图;
图2是根据本公开的示例性配置界面的示意图;
图3是根据本公开的第二终端显示通知信息的示例性示意图
图4是根据本公开的第一终端展示听写进度信息的示例性示意图;
图5是本公开的第一终端对听写图像信息进行批改的示例性示意图;
图6是根据本公开的第二终端展示批改结果的示例性示意图;
图7是根据本公开的展示听写项的解释信息的示例性示意图;
图8是根据本公开的听写交互装置的一个实施例的结构示意图;
图9是本公开的一个实施例的听写交互方法可以应用于其中的示例性系统架构;
图10是根据本公开实施例提供的电子设备的基本结构的示意图。
下面将参照附图更详细地描述本公开的实施例。虽然附图中显示了本公开的某些实施例,然而应当理解的是,本公开可以通过各种形式来实现,而且不应该被解释为限于这里阐述的实施例,相反提供这些实施例是为了更加透彻和完整地理解本公开。应当理解的是,本公开的附图及实施例仅用于示例性作用,并非用于限制本公开的保护范围。
应当理解,本公开的方法实施方式中记载的各个步骤可以按照不同的顺序执行,和/或并行执行。此外,方法实施方式可以包括附加的步骤和/或省略执行示出的步骤。本公开的范围在此方面不受限制。
本文使用的术语“包括”及其变形是开放性包括,即“包括但不限于”。术语“基于”是“至少部分地基于”。术语“一个实施例”表示“至少一个实施例”;术语“另一实施例”表示“至少一个另外的实施例”;术语“一些实施例”表示“至少一些实施例”。其他术语的相关 定义将在下文描述中给出。
需要注意,本公开中提及的“第一”、“第二”等概念仅用于对不同的装置、模块或单元进行区分,并非用于限定这些装置、模块或单元所执行的功能的顺序或者相互依存关系。
需要注意,本公开中提及的“一个”、“多个”的修饰是示意性而非限制性的,本领域技术人员应当理解,除非在上下文另有明确指出,否则应该理解为“一个或多个”。
本公开实施方式中的多个装置之间所交互的消息或者信息的名称仅用于说明性的目的,而并不是用于对这些消息或信息的范围进行限制。
请参考图1,其示出了根据本公开的听写交互方法的一个实施例的流程。如图1所示该听写交互方法,包括以下步骤:
步骤101,展示配置界面,以及通过配置界面确定目标听写材料和听写参与用户。
在本实施例中,听写交互方法的执行主体(例如第一终端)可以展示配置界面。
在本实施例中,上述配置界面可以用于供听写发起用户配置目标听写材料和听写参与用户标识。
在这里,听写发起用户可以是发起听写的用户,换句话说,可以是布置听写任务的用户。
在这里,目标听写材料可以是待进行听写的内容。目标听写材料的具体形式不做限定,作为示例,目标听写材料可以是字、词、句子、段落等。
在这里,听写参与用户标识可以指示听取目标听写材料进行写操作的用户。听写参与用户标识可以是一个,也可以是至少两个,在此不做限定。
作为示例,听写发起用户可以是教师,听写参与用户可以是学生。
在本实施例中,上述配置界面可以是一个界面,也可以是多个界面。换句话说,配置目标听写材料与配置听写参与用户标识, 可以在同一界面中进行,也可以在不同界面中进行。如果在不同界面中进行,这些不同的界面可以统称为配置界面。
步骤102,将目标听写材料对应的语音文件,发送至听写参与用户标识对应的终端设备。
在本实施例中,上述执行主体可以将上述目标听写材料对应的语音文件,发送至听写参与用户标识对应的终端设备。
在本实施例中,听写参与用户标识指示的听写参与用户,可以基于上述语音文件进行听写。
在一些应用场景中,上述配置界面可以设置听写发起确认控件,上述听写发起用户可以在配置目标听写材料和听写参与用户标识之后,触发听写发起确认控件。然后,上述执行主体可以基于听写发起用户的配置,向听写参与用户对应的终端设备,发送目标听写材料对应的语音文件。
在一些应用场景中,发送上述语音文件,可以是上述执行主体直接发送的,还可以是上述执行主体间接发送的。在一些实现方式中,上述执行主体可以将目标听写材料发送至服务器,服务器可以获取已生成的语音文件或者根据目标听写材料生成语音文件。
在这里,听写参与用户标识对应的终端设备(本申请中可以称为第二终端设备),可以播放语音文件。听写参与用户听着终端设备所播放的语音,进行听写。
在一些可选的实现方式中,上述第二终端设备,可以在接收到语音文件即开始播放,也可以由听写参与用户选择播放时间再进行播放。
请参考图2,其示出了配置界面的示例性示意图。在图2中,可以选择课文标题为“古诗两首”的课文,然后在选择“古诗两首”的识字表表,则自动呈现“绿”“丝”“绦”三个听写项。
需要说明的是,本实施例提供的听写交互方法,通过先展示配置界面,供听写发起用户配置目标听写材料和听写参与用户标识;然后,将目标听写材料对应的语音文件,发送至听写参与用 户。由此,听写发起用户可以在线上布置听写任务,听写过程不受空间的限制。
在一些实施例中,上述配置界面可以包括以下至少一项但不限于:候选听写内容标识和听写内容补充区域。
在这里,候选听写内容标识可以指示预先存储的候选听写内容。
在一些应用场景中,上述听写发起用户可以通过选取候选听写内容标识,确定目标听写材料。可选的,上述执行主体可以响应于检测到针对候选听写内容标识的第一选取操作,将第一选取操作针对的候选听写标识所指示的候选听写内容,确定为目标听写材料。
在这里,第一选取操作的具体内容,在此不做限定。
在一些应用场景中,上述听写发起用户,可以通过实施一系列的选择操作,实现配置目标听写材料。
在一些应用场景中,可以上述听写发起用户,可以通过依次选择学科、教材、课文。听写发起用户选择了课文标识,可以将课文标识预先关联的词汇,作为候选听写内容,供用户候选听写内容中选取目标听写材料。
请参考图2,“绿”“丝”“绦”三个听写项标识;理解为候选听写内容标识。听写发起用户可以从“绿”“丝”“绦”三个听写项中选取目标听写材料。
需要说明的是,通过预先设置候选听写内容,以及展示候选听写内容标识,可以便于听写发起用户布置听写任务。
在一些实施例中,上述配置界面包括听写内容补充区域。上述步骤101,可以包括:响应于获取到在听写内容补充区域输入的听写补充内容,将听写补充内容确定为目标听写材料。
在一些应用场景中,上述听写发起用户可以在听写内容补充区域输入期望的听写材料,然后触发确认控件,以确认将输入的听写材料确定为最终输入到上述听写内容补充区域的材料。上述执行主体可以将听写内容补充区域输入的听写材料也作为目标听 写材料。
需要说明的是,通过设置听写内容补充区域,可以使得听写发起用户灵活布置听写任务。
请参考图2,图2中的“选择听写内容”下方的框可以理解为听写内容补充区域。听写内容补充区域中的“剪刀”可以是听写补充内容。
在一些实施例中,听写方式信息可以指示听写方式。
在一些实施例中,听写方式信息可以包括但是不限于以下至少一项:重复读词次数、读词间隔时长、听写顺序指示信息。
在这里,重复读词次数可以指示单个听写项的重复播放次数。
在这里,读词间隔时长可以指示播放词语结束到开始播放该听写项的下一听写项之间的时长。
在这里,听写顺序指示信息,可以指示听写材料中听写项的播放顺序。作为示例,播放顺序可以是随机播放,也可以按顺序播放。
在一些应用场景中,如果听写发起用户设置乱序播放,在所有的听写参与用户处,可以采用同一种乱序方式。听写发起用户接收到的多份听写内容图像中听写项的顺序是一致的,从而可以提高听写发起用户的批改效率。
在这里,上述方法可以包括:通过所述配置界面,获取听写发起用户针对所述目标听写材料配置的听写方式信息。
在这里,所述终端设备可以采用听写方式信息指示的听写方式,播放所述语音文件。
可以理解,不同的听写方式信息,可能使得听写任务的难度发生变化。作为示例,重复读词次数越大,听写难度越低;读词间隔时长越大,听写难度越低;顺序听写比乱序听写,听写难度低。
需要说明的是,通过设置听写方式信息,可以使得听写发起用户针对本次听写任务灵活设置符合实际应用的听写方式信息,即灵活设置听写任务的难度。
请参考图2,其示出了配置听写方式信息的场景示意图。读词间隔、听写顺序、读词次数均可以作为听写方式信息进行配置。
在一些实施例中,目标听写材料包括至少一个听写项。听写项可以是字、词、句子或者段落。目标听写材料对应的语音文件可以通过以下方式确定:针对目标听写材料中的各个听写项,确定是否已生成该听写项对应的语音文件;响应于确定未生成,合成该听写项对应的语音文件。然后,可以将合成的语音文件存入到目标听写材料对应的语音文件中。
需要说明的是,上述确定目标听写材料对应的语音文件的步骤,可以是上述执行主体执行的,也可以是上述支持上述执行主体的服务器执行的。
需要说明的是,通过检测各个听写项是否具有对应的语言文件,如果没有则合成语音文件,可以降低听写发起用户的工作量,减少布置听写任务的步骤,提高布置听写任务的效率。
在一些实施例中,上述第二终端(听写参与用户标识对应的终端设备)可以接收听写通知,并且可以展示听写通知。上述听写通知用于通知听写参与用户相应的听写任务。
在一些应用场景中,请参考图3,图3示出了终端展示的听写通知。在图3中,听写通知中可以包括“老师给你布置了听写任务”和“《古诗两首》等课文需要听写,赶紧来完成吧”;并且,听写通知中可以包括确认控件(图3中标示“立即查看”字样)。
在一些实施例中,第二终端可以展示听写通知,以及响应于检测到预定义的听写开始操作,开启听写进程。
在这里,预定义的听写开始操作的具体内容,可以根据实际应用场景设置,在此不做限定。
在一些实施例中,上述第二终端的听写开启时机,可以是接收到上述听写通知的时候,也可以是上述听写发起用户配置的听写开启时间,也可以是上述听写参与用户自己选择的时间。
作为示例,预定义的听写开始操作,可以包括对听写通知的触发操作。换句话说,触发听写通知,即可开始听写进程。
作为示例,触发听写通知之后,可以展示听写开始确认控件。听写参与用户触发听写开始确认控件,可以作为预定义的听写开始操作。
作为示例,应用中可以设置听写任务查看入口,用户从听写任务查看入口,可以看到自己的听写任务。未完成的听写任务可以关联设置听写开始确认控件。听写参与用户对听写开始确认控件的触发操作,可以作为预定义的听写开始操作。换句话说,听写参与用户对未完成的听写任务进行触发,可以开启该听写任务。
在一些实施例中,上述第二执行主体响应于预定义的听写开始操作,可以播放语音文件,以及采集听写参与用户的用户图像。
在这里,听写参与用户的用户图像,可以包括听写参与用户在听写过程中的任意人体部位的图像。作为示例,可以采集听写参与用户的上半身图像,也可以采集听写参与用户进行书写的手部图像。
在一些实施例中,上述听写参与用户可以在第二终端上书写,也可以采用纸笔书写。上述第一执行主体可以包括摄像头,摄像头可以直接对准听写参与用户的手写位置进行图像采集,还可以通过光路设置(例如设置反光镜)等,对听写参与用户的手写位置进行图像采集。即第二终端可以通过摄像头接收听写参与用户的书写图像。
在一些实施例中,上述第二终端还可以通过显示屏接收听写参与用户的书写图像。
在一些实施例中,如果上述听写参与用户在第二终端上书写,上述第二终端可以通过录屏,获取听写参与用户的书写过程。
需要说明的是,通过第二终端接收用户在显示屏上输入的书写内容,可以避免用户准备纸笔进行听写的步骤,降低听写工具对实施听写的限制。从而,可以使得听写参与用户可以随时随地开始听写,提高听写的效率。
在一些实施例中,上述第二终端可以将听写参与用户的用户图像,实时传输到预设的电子设备(例如服务器)。并且,第二终 端可以将听写进度信息,实时反馈至第一终端。如果第一终端检测到获取用户图像的操作,第一终端可以从上述预设的电子设备拉流,展示用户图像(或视频)。
在一些实施例中,上述第二终端可以响应于确定听写结束,采集听写内容图像。
在这里,确定听写结束的具体实现方式,可以根据实际应用场景设置,在此不做限定。
作为示例,上述第二终端可以响应于检测到针对预设听写结束控件的触发操作,确定听写结束。
作为示例,上述第二终端可以响应于从听写开始为起点经过了预设听写时长,确定听写结束。预设听写时长可以是根据目标听写材料,预先设置的。
作为示例,上述第二终端可以响应于目标听写材料播放完成为七点经过了预设时长,确定听写结束。作为示例,目标听写材料播放完成5秒,可以确定听写结束。
在这里,上述第二终端可以响应于听写结束,吊起图像采集功能。
在这里,上述第二终端可以展示引导信息,例如“请将听写内容放在屏幕中”,以引导听写参与用户拍摄听写内容图像。
在一些实施例中,如果上述听写参与用户在第二终端上书写,上述第二终端可以获取听写参与用户的书写听写内容图像。
需要说明的是,第二终端响应于确定听写结束,而所采集的听写内容图像,相对于听写过程中采集的图像,一般更为清晰。听写结束采集的听写内容图像作为听写发起用户的批改基础,可以避免因图像不清晰而造成的批改效率降低、批改不准确。
在一些应用场景中,可以以班级为单位,展示跟该班级相关的各个听写任务。作为示例,可以展示三年一班的语文听写任务和英语听写任务。
在一些应用场景中,可以以听写发起用户为单位,展示该听 写发起用户所发起的各个听写任务。作为示例,对于李老师,可以展示李老师针对三年一班发起的语文听写任务,以及展示李老师针对三年二班发起的语文听写任务。
在一些实施例中,可以展示听写任务的整体执行进度信息。整体执行进度信息可以包括以下至少一项但不限于:听写任务的听写参与用户总数量、完成听写任务的听写参与用户数量、未完成听写任务的听写参与用户数量。
在一些实施例中,上述执行主体可以展示各个听写参与用户的听写进度信息。
在一些应用场景中,请参考图4,其示出了执行主体展示各个听写参与用于的听写进度信息的应用场景。在图4中,对于听写参与用户“张三”,可以对应展示听写进度信息“听写完成”;对于听写参与用户“李四”,可以对应展示听写进度信息“听写20%”;对于听写参与用户“王五”,可以对应展示听写进度信息“已完成”;对于听写参与用户“宋六”,可以对应展示听写进度信息“未开始”。
需要说明的是,通过展示听写参与用户的听写进度信息,可以使得听写发起用户及时获取所布置的听写任务的执行情况,由此,听写发起用户可以以此为基础,及时对未开始的听写参与用户进行提醒,提高听写交互的效率。
在一些实施例中,上述执行主体可以对于听写进行中或者听写完成的听写参与用户,获取以及播放听写过程中的图像。换句话说,对进行中或者已完成的听写参与用户,听写发起用户可以查看听写过程的用户图像或者视频。
需要说明的是,通过第二终端的拍照或者录像功能,可以记录听写过程中的听写参与用户的用户图像。从而,听写发起用户查看听写过程中的用户图像,可以使得听写发起用户对于听写过程进行实时或者非实时的监督,提高听写效率。
在一些实施例中,相同时间点采集的用户图像和书写图像之间具有关联关系;以及所述方法包括:展示具有关联关系的用户图像和书写图像。
可选的,书写图像可以是用户通过纸笔手写,第二终端采集的;可选的,书写图像还可以是用户通过第二终端的显示屏手写的,第二终端通过录屏等方式获取的。
需要说明的是,通过展示具有关联关系的用户图像和书写图像,可以向听写发起用户展示实际听写过程。从而,听写发起用户可以通过实际听写过程中的书写图像和用户图像,获取更多实际听写过程中的用户信息,从而对听写过程进行有效监督。
在一些实施例中,上述执行主体可以展示听写内容图像和目标听写材料;根据对目标听写材料中的第二选取操作,生成听写内容图像的批改结果。
需要说明的是,并列展示听写内容图像和听写材料,可以方便作业批改用户(通常为听写发起用户)对听写出错部分的选择。
请参考图5,其示出了批改过程的场景示意图。在图5中,展示了张三的听写内容图像,并且展示了目标听写材料。张三的作业中,将“绦”错写成了“条”,由此,作用批改用户可以在目标听写材料中选取“绦”(以阴影进行表示),由此可以得到作业批改结果。
在一些实施例中,所述方法还包括将批改结果返回给目标待批改用户,其中,批改结果包括目标听写材料和错误项指示信息。
在这里,错误项指示信息可以指示听写内容图像中的错误项。
需要说明的是,在线上以目标听写材料为基础,反馈批改结果,可以将批改结果实时反馈给听写参与用户,提升了反馈效率和听写参与用户的学习效果。
在这里,第二终端响应于检测到针对听写项的触发操作,获取触发操作所针对的听写项的解释信息,展示所获取的解释信息。
在这里,解释信息可以用于解释听写项。解释信息可以预先存储。作为示例,解释信息的来源可以包括但是不限于以下至少一项:字典、教科书等。
请参考图6,其示出了展示第二终端展示批改结果的场景示意图,在图6中,可以展示张三(即第二终端的登录用户)本人的 答案,即“我的答案”。并且展示批改结果,批改结果可以采用各种形式,作为示例,批改结果可以采用在目标听写材料的听写项上附加错误项指示信息的形式展示。
请参考图7,其示出了第二终端展示解释信息的场景示意图。张三可以点击图6中阴影显示的“绦”,然后,第二终端可以展示图6所示的“绦”的解释信息。
需要说明的是,以目标听写材料为基础,将目标听写材料中的听写项与解释信息相关联,可以使得听写参与用户快速获取听写项的详细信息,从而能够及时有效地巩固听写参与用户的学习效果。
进一步参考图8,作为对上述各图所示方法的实现,本公开提供了一种听写交互装置的一个实施例,该装置实施例与图1所示的方法实施例相对应,该装置具体可以应用于各种电子设备中。
如图8所示,本实施例的听写交互装置包括:展示单元801、和发送单元。其中,展示单元,用于展示配置界面,以及通过所述配置界面确定目标听写材料和听写参与用户,其中,所述配置界面用于供听写发起用户配置目标听写材料和听写参与用户标识;发送单元,用于将所述目标听写材料对应的语音文件,发送至听写参与用户标识对应的第二终端,其中,听写参与用户基于所述语音文件进行听写。
在本实施例中,听写交互装置的展示单元801、和发送单元的具体处理及其所带来的技术效果可分别参考图8对应实施例中步骤101和步骤102的相关说明,在此不再赘述。
在一些实施例中,所述配置界面包括以下至少一项:候选听写内容标识和听写内容补充区域;以及所述展示配置界面,以及通过所述配置界面确定目标听写材料和听写参与用户,包括:展示候选听写内容标识,以及响应于检测针对候选听写内容标识的第一选取操作,将第一选取操作针对的候选听写标识所指示的候 选听写内容,确定为所述目标听写材料;展示听写内容补充区域,以及响应于获取到在听写内容补充区域输入的听写补充内容,将听写补充内容确定为目标听写材料。
在一些实施例中,所述配置界面包括听写方式信息配置区域;以及所述装置还用于:通过所述配置界面,获取听写发起用户针对所述目标听写材料配置的听写方式信息,其中,所述终端设备采用听写方式信息指示的听写方式播放所述语音文件,所述听写方式信息包括以下至少一项:重复读词次数、读词间隔时长、听写顺序指示信息。
在一些实施例中,所述目标听写材料包括至少一个听写项,其中,目标听写材料对应的语音文件通过以下方式确定:针对目标听写材料中的各个听写项,确定是否已生成该听写项对应的语音文件;响应于确定未生成,合成该听写项对应的语音文件。
在一些实施例中,所述第二终端展示听写通知。
在一些实施例中,所述第二终端响应于预定义的听写开始操作,播放所述语音文件,以及采集听写参与用户的用户图像。
在一些实施例中,所述第二终端通过显示屏和/或摄像头接收听写参与用户的书写图像。
在一些实施例中,第二终端响应于确定听写结束,采集听写内容图像。
在一些实施例中,所述装置还用于:展示各个听写参与用户的听写进度信息。
在一些实施例中,所述装置还用于:针对听写进行中或者听写完成的听写参与用户,展示听写过程中的用户图像。
在一些实施例中,相同时间点采集的用户图像和书写图像之间具有关联关系;以及所述装置还用于:展示具有关联关系的用户图像和书写图像。
在一些实施例中,所述装置还用于:展示听写内容图像和目标听写材料;根据对目标听写材料中听写项的第二选取操作,生成听写内容图像的批改结果,其中,批改结果包括目标听写材料 和错误项指示信息。
在一些实施例中,第二终端展示批改结果,以及响应于检测到针对听写项的触发操作,获取以及展示触发操作所针对的听写项的解释信息。
请参考图9,图9示出了本公开的一个实施例的听写交互方法可以应用于其中的示例性系统架构。
如图9所示,系统架构可以包括终端设备901、902、903,网络904,服务器905。网络904用以在终端设备901、902、903和服务器905之间提供通信链路的介质。网络904可以包括各种连接类型,例如有线、无线通信链路或者光纤电缆等等。
终端设备901、902、903可以通过网络904与服务器905交互,以接收或发送消息等。终端设备901、902、903上可以安装有各种客户端应用,例如网页浏览器应用、搜索类应用、新闻资讯类应用。终端设备901、902、903中的客户端应用可以接收用户的指令,并根据用户的指令完成相应的功能,例如根据用户的指令在信息中添加相应信息。
终端设备901、902、903可以是硬件,也可以是软件。当终端设备901、902、903为硬件时,可以是具有显示屏并且支持网页浏览的各种电子设备,包括但不限于智能手机、平板电脑、电子书阅读器、MP3播放器(Moving Picture Experts Group Audio Layer III,动态影像专家压缩标准音频层面3)、MP4(Moving Picture Experts Group Audio Layer IV,动态影像专家压缩标准音频层面4)播放器、膝上型便携计算机和台式计算机等等。当终端设备901、902、903为软件时,可以安装在上述所列举的电子设备中。其可以实现成多个软件或软件模块(例如用来提供分布式服务的软件或软件模块),也可以实现成单个软件或软件模块。在此不做具体限定。
服务器905可以是提供各种服务的服务器,例如接收终端设备901、902、903发送的信息获取请求,根据信息获取请求通过 各种方式获取信息获取请求对应的展示信息。并展示信息的相关数据发送给终端设备901、902、903。
需要说明的是,本公开实施例所提供的听写交互方法可以由终端设备执行,相应地,听写交互装置可以设置在终端设备901、902、903中。此外,本公开实施例所提供的听写交互方法还可以由终端设备和服务器905联合执行,相应地,听写交互装置可以设置于终端设备和服务器905中。
应该理解,图9中的终端设备、网络和服务器的数目仅仅是示意性的。根据实现需要,可以具有任意数目的终端设备、网络和服务器。
下面参考图10,其示出了适于用来实现本公开实施例的电子设备(例如图9中的终端设备或服务器)的结构示意图。本公开实施例中的终端设备可以包括但不限于诸如移动电话、笔记本电脑、数字广播接收器、PDA(个人数字助理)、PAD(平板电脑)、PMP(便携式多媒体播放器)、车载终端(例如车载导航终端)等等的移动终端以及诸如数字TV、台式计算机等等的固定终端。图10示出的电子设备仅仅是一个示例,不应对本公开实施例的功能和使用范围带来任何限制。
如图10所示,电子设备可以包括处理装置(例如中央处理器、图形处理器等)1001,其可以根据存储在只读存储器(ROM)1002中的程序或者从存储装置1008加载到随机访问存储器(RAM)1003中的程序而执行各种适当的动作和处理。在RAM 1003中,还存储有电子设备1000操作所需的各种程序和数据。处理装置1001、ROM 1002以及RAM 1003通过总线1004彼此相连。输入/输出(I/O)接口1005也连接至总线1004。
通常,以下装置可以连接至I/O接口1005:包括例如触摸屏、触摸板、键盘、鼠标、摄像头、麦克风、加速度计、陀螺仪等的输入装置1006;包括例如液晶显示器(LCD)、扬声器、振动器等的输出装置1007;包括例如磁带、硬盘等的存储装置1008;以及 通信装置1009。通信装置1009可以允许电子设备与其他设备进行无线或有线通信以交换数据。虽然图10示出了具有各种装置的电子设备,但是应理解的是,并不要求实施或具备所有示出的装置。可以替代地实施或具备更多或更少的装置。
特别地,根据本公开的实施例,上文参考流程图描述的过程可以被实现为计算机软件程序。例如,本公开的实施例包括一种计算机程序产品,其包括承载在非暂态计算机可读介质上的计算机程序,该计算机程序包含用于执行流程图所示的方法的程序代码。在这样的实施例中,该计算机程序可以通过通信装置1009从网络上被下载和安装,或者从存储装置1008被安装,或者从ROM 1002被安装。在该计算机程序被处理装置1001执行时,执行本公开实施例的方法中限定的上述功能。
需要说明的是,本公开上述的计算机可读介质可以是计算机可读信号介质或者计算机可读存储介质或者是上述两者的任意组合。计算机可读存储介质例如可以是——但不限于——电、磁、光、电磁、红外线、或半导体的系统、装置或器件,或者任意以上的组合。计算机可读存储介质的更具体的例子可以包括但不限于:具有一个或多个导线的电连接、便携式计算机磁盘、硬盘、随机访问存储器(RAM)、只读存储器(ROM)、可擦式可编程只读存储器(EPROM或闪存)、光纤、便携式紧凑磁盘只读存储器(CD-ROM)、光存储器件、磁存储器件、或者上述的任意合适的组合。在本公开中,计算机可读存储介质可以是任何包含或存储程序的有形介质,该程序可以被指令执行系统、装置或者器件使用或者与其结合使用。而在本公开中,计算机可读信号介质可以包括在基带中或者作为载波一部分传播的数据信号,其中承载了计算机可读的程序代码。这种传播的数据信号可以采用多种形式,包括但不限于电磁信号、光信号或上述的任意合适的组合。计算机可读信号介质还可以是计算机可读存储介质以外的任何计算机可读介质,该计算机可读信号介质可以发送、传播或者传输用于由指令执行系统、装置或者器件使用或者与其结合使用的程序。 计算机可读介质上包含的程序代码可以用任何适当的介质传输,包括但不限于:电线、光缆、RF(射频)等等,或者上述的任意合适的组合。
在一些实施方式中,客户端、服务器可以利用诸如HTTP(HyperText Transfer Protocol,超文本传输协议)之类的任何当前已知或未来研发的网络协议进行通信,并且可以与任意形式或介质的数字数据通信(例如,通信网络)互连。通信网络的示例包括局域网(“LAN”),广域网(“WAN”),网际网(例如,互联网)以及端对端网络(例如,ad hoc端对端网络),以及任何当前已知或未来研发的网络。
上述计算机可读介质可以是上述电子设备中所包含的;也可以是单独存在,而未装配入该电子设备中。
上述计算机可读介质承载有一个或者多个程序,当上述一个或者多个程序被该电子设备执行时,使得该电子设备:展示配置界面,以及通过所述配置界面确定目标听写材料和听写参与用户,其中,所述配置界面用于供听写发起用户配置目标听写材料和听写参与用户标识;将所述目标听写材料对应的语音文件,发送至听写参与用户标识对应的第二终端,其中,听写参与用户基于所述语音文件进行听写。
可以以一种或多种程序设计语言或其组合来编写用于执行本公开的操作的计算机程序代码,上述程序设计语言包括但不限于面向对象的程序设计语言—诸如Java、Smalltalk、C++,还包括常规的过程式程序设计语言—诸如“C”语言或类似的程序设计语言。程序代码可以完全地在用户计算机上执行、部分地在用户计算机上执行、作为一个独立的软件包执行、部分在用户计算机上部分在远程计算机上执行、或者完全在远程计算机或服务器上执行。在涉及远程计算机的情形中,远程计算机可以通过任意种类的网络——包括局域网(LAN)或广域网(WAN)—连接到用户计算机,或者,可以连接到外部计算机(例如利用因特网服务提供商来通过因特网连接)。
附图中的流程图和框图,图示了按照本公开各种实施例的系统、方法和计算机程序产品的可能实现的体系架构、功能和操作。在这点上,流程图或框图中的每个方框可以代表一个模块、程序段、或代码的一部分,该模块、程序段、或代码的一部分包含一个或多个用于实现规定的逻辑功能的可执行指令。也应当注意,在有些作为替换的实现中,方框中所标注的功能也可以以不同于附图中所标注的顺序发生。例如,两个接连地表示的方框实际上可以基本并行地执行,它们有时也可以按相反的顺序执行,这依所涉及的功能而定。也要注意的是,框图和/或流程图中的每个方框、以及框图和/或流程图中的方框的组合,可以用执行规定的功能或操作的专用的基于硬件的系统来实现,或者可以用专用硬件与计算机指令的组合来实现。
描述于本公开实施例中所涉及到的单元可以通过软件的方式实现,也可以通过硬件的方式来实现。其中,单元的名称在某种情况下并不构成对该单元本身的限定,例如,展示单元还可以被描述为“展示配置界面的单元”。
本文中以上描述的功能可以至少部分地由一个或多个硬件逻辑部件来执行。例如,非限制性地,可以使用的示范类型的硬件逻辑部件包括:现场可编程门阵列(FPGA)、专用集成电路(ASIC)、专用标准产品(ASSP)、片上系统(SOC)、复杂可编程逻辑设备(CPLD)等等。
在本公开的上下文中,机器可读介质可以是有形的介质,其可以包含或存储以供指令执行系统、装置或设备使用或与指令执行系统、装置或设备结合地使用的程序。机器可读介质可以是机器可读信号介质或机器可读储存介质。机器可读介质可以包括但不限于电子的、磁性的、光学的、电磁的、红外的、或半导体系统、装置或设备,或者上述内容的任何合适组合。机器可读存储介质的更具体示例会包括基于一个或多个线的电气连接、便携式计算机盘、硬盘、随机存取存储器(RAM)、只读存储器(ROM)、可擦除可编程只读存储器(EPROM或快闪存储器)、光纤、便捷 式紧凑盘只读存储器(CD-ROM)、光学储存设备、磁储存设备、或上述内容的任何合适组合。
以上描述仅为本公开的较佳实施例以及对所运用技术原理的说明。本领域技术人员应当理解,本公开中所涉及的公开范围,并不限于上述技术特征的特定组合而成的技术方案,同时也应涵盖在不脱离上述公开构思的情况下,由上述技术特征或其等同特征进行任意组合而形成的其它技术方案。例如上述特征与本公开中公开的(但不限于)具有类似功能的技术特征进行互相替换而形成的技术方案。
此外,虽然采用特定次序描绘了各操作,但是这不应当理解为要求这些操作以所示出的特定次序或以顺序次序执行来执行。在一定环境下,多任务和并行处理可能是有利的。同样地,虽然在上面论述中包含了若干具体实现细节,但是这些不应当被解释为对本公开的范围的限制。在单独的实施例的上下文中描述的某些特征还可以组合地实现在单个实施例中。相反地,在单个实施例的上下文中描述的各种特征也可以单独地或以任何合适的子组合的方式实现在多个实施例中。
尽管已经采用特定于结构特征和/或方法逻辑动作的语言描述了本主题,但是应当理解所附权利要求书中所限定的主题未必局限于上面描述的特定特征或动作。相反,上面所描述的特定特征和动作仅仅是实现权利要求书的示例形式。
Claims (16)
- 一种听写交互方法,其特征在于,应用于第一终端,所述方法包括:展示配置界面,以及通过所述配置界面确定目标听写材料和听写参与用户,其中,所述配置界面用于供听写发起用户配置目标听写材料和听写参与用户标识;将所述目标听写材料对应的语音文件,发送至听写参与用户标识对应的第二终端,其中,听写参与用户基于所述语音文件进行听写。
- 根据权利要求1所述的方法,其特征在于,所述配置界面包括以下至少一项:候选听写内容标识和听写内容补充区域;以及所述展示配置界面,以及通过所述配置界面确定目标听写材料和听写参与用户,包括:展示候选听写内容标识,以及响应于检测针对候选听写内容标识的第一选取操作,将第一选取操作针对的候选听写标识所指示的候选听写内容,确定为所述目标听写材料;展示听写内容补充区域,以及响应于获取到在听写内容补充区域输入的听写补充内容,将听写补充内容确定为目标听写材料。
- 根据权利要求1所述的方法,其特征在于,所述配置界面包括听写方式信息配置区域;以及所述方法还包括:通过所述配置界面,获取听写发起用户针对所述目标听写材料配置的听写方式信息,其中,所述终端设备采用听写方式信息指示的听写方式播放所述语音文件,所述听写方式信息包括以下至少一项:重复读词次数、读词间隔时长、听写顺序指示信息。
- 根据权利要求1所述的方法,其特征在于,所述目标听写 材料包括至少一个听写项,其中,目标听写材料对应的语音文件通过以下方式确定:针对目标听写材料中的各个听写项,确定是否已生成该听写项对应的语音文件;响应于确定未生成,合成该听写项对应的语音文件。
- 根据权利要求1所述的方法,其特征在于,所述第二终端展示听写通知。
- 根据权利要求1所述的方法,其特征在于,所述第二终端响应于预定义的听写开始操作,播放所述语音文件,以及采集听写参与用户的用户图像。
- 根据权利要求6所述的方法,其特征在于,所述第二终端通过显示屏和/或摄像头接收听写参与用户的书写图像。
- 根据权利要求6所述的方法,其特征在于,第二终端响应于确定听写结束,采集听写内容图像。
- 根据权利要求1所述的方法,其特征在于,所述方法包括:展示各个听写参与用户的听写进度信息。
- 根据权利要求1所述的方法,其特征在于,所述方法包括:针对听写进行中或者听写完成的听写参与用户,展示听写过程中的用户图像。
- 根据权利要求10所述的方法,其特征在于,相同时间点采集的用户图像和书写图像之间具有关联关系;以及所述方法包括:展示具有关联关系的用户图像和书写图像。
- 根据权利要求8所述的方法,其特征在于,所述方法还包括:展示听写内容图像和目标听写材料;根据对目标听写材料中听写项的第二选取操作,生成听写内容图像的批改结果,其中,批改结果包括目标听写材料和错误项指示信息。
- 根据权利要求12所述的方法,其特征在于,第二终端展示批改结果,以及响应于检测到针对听写项的触发操作,获取以及展示触发操作所针对的听写项的解释信息。
- 一种听写交互装置,其特征在于,应用于第一终端,所述装置包括:展示单元,用于展示配置界面,以及通过所述配置界面确定目标听写材料和听写参与用户,其中,所述配置界面用于供听写发起用户配置目标听写材料和听写参与用户标识;发送单元,用于将所述目标听写材料对应的语音文件,发送至听写参与用户标识对应的第二终端,其中,听写参与用户基于所述语音文件进行听写。
- 一种电子设备,其特征在于,包括:一个或多个处理器;存储装置,用于存储一个或多个程序,当所述一个或多个程序被所述一个或多个处理器执行,使得所述一个或多个处理器实现如权利要求1-13中任一所述的方法。
- 一种计算机可读介质,其上存储有计算机程序,其特征在于,该程序被处理器执行时实现如权利要求1-13中任一所述的 方法。
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| CN202010714796.2A CN111930453A (zh) | 2020-07-21 | 2020-07-21 | 听写交互方法、装置和电子设备 |
| CN202010714796.2 | 2020-07-21 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2022017203A1 true WO2022017203A1 (zh) | 2022-01-27 |
Family
ID=73315319
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/CN2021/105611 Ceased WO2022017203A1 (zh) | 2020-07-21 | 2021-07-09 | 听写交互方法、装置和电子设备 |
Country Status (2)
| Country | Link |
|---|---|
| CN (1) | CN111930453A (zh) |
| WO (1) | WO2022017203A1 (zh) |
Families Citing this family (3)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN111930453A (zh) * | 2020-07-21 | 2020-11-13 | 北京字节跳动网络技术有限公司 | 听写交互方法、装置和电子设备 |
| CN112712737A (zh) * | 2021-01-13 | 2021-04-27 | 百度在线网络技术(北京)有限公司 | 交互方法、装置、设备以及存储介质 |
| CN112817558A (zh) * | 2021-02-19 | 2021-05-18 | 北京大米科技有限公司 | 听写数据处理的方法、装置、可读存储介质和电子设备 |
Citations (6)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| KR20010079164A (ko) * | 2001-06-19 | 2001-08-22 | 김성기 | 따라쓰기 학습을 위한 문자인식 방법 및 장치 |
| CN102419918A (zh) * | 2010-12-30 | 2012-04-18 | 深圳市高德讯科技有限公司 | 一种教师家庭作业布置的方法及其系统、学生家庭作业系统 |
| CN110309350A (zh) * | 2018-03-21 | 2019-10-08 | 腾讯科技(深圳)有限公司 | 背诵任务的处理方法、系统、装置、介质及电子设备 |
| CN110490780A (zh) * | 2019-08-27 | 2019-11-22 | 北京赢裕科技有限公司 | 一种辅助语文学习的方法及系统 |
| CN111081117A (zh) * | 2019-05-10 | 2020-04-28 | 广东小天才科技有限公司 | 一种书写检测方法及电子设备 |
| CN111930453A (zh) * | 2020-07-21 | 2020-11-13 | 北京字节跳动网络技术有限公司 | 听写交互方法、装置和电子设备 |
Family Cites Families (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US9741342B2 (en) * | 2014-11-26 | 2017-08-22 | Panasonic Intellectual Property Corporation Of America | Method and apparatus for recognizing speech by lip reading |
| CN105702246A (zh) * | 2016-03-17 | 2016-06-22 | 广东小天才科技有限公司 | 一种辅助用户进行听写的方法及装置 |
| CN110263334A (zh) * | 2019-06-06 | 2019-09-20 | 深圳市柯达科电子科技有限公司 | 一种辅助外语学习的方法和可读存储介质 |
| CN111079423A (zh) * | 2019-08-02 | 2020-04-28 | 广东小天才科技有限公司 | 一种听写报读音频的生成方法、电子设备及存储介质 |
-
2020
- 2020-07-21 CN CN202010714796.2A patent/CN111930453A/zh active Pending
-
2021
- 2021-07-09 WO PCT/CN2021/105611 patent/WO2022017203A1/zh not_active Ceased
Patent Citations (6)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| KR20010079164A (ko) * | 2001-06-19 | 2001-08-22 | 김성기 | 따라쓰기 학습을 위한 문자인식 방법 및 장치 |
| CN102419918A (zh) * | 2010-12-30 | 2012-04-18 | 深圳市高德讯科技有限公司 | 一种教师家庭作业布置的方法及其系统、学生家庭作业系统 |
| CN110309350A (zh) * | 2018-03-21 | 2019-10-08 | 腾讯科技(深圳)有限公司 | 背诵任务的处理方法、系统、装置、介质及电子设备 |
| CN111081117A (zh) * | 2019-05-10 | 2020-04-28 | 广东小天才科技有限公司 | 一种书写检测方法及电子设备 |
| CN110490780A (zh) * | 2019-08-27 | 2019-11-22 | 北京赢裕科技有限公司 | 一种辅助语文学习的方法及系统 |
| CN111930453A (zh) * | 2020-07-21 | 2020-11-13 | 北京字节跳动网络技术有限公司 | 听写交互方法、装置和电子设备 |
Also Published As
| Publication number | Publication date |
|---|---|
| CN111930453A (zh) | 2020-11-13 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| EP4343514A1 (en) | Display method and apparatus, and device and storage medium | |
| CN111970577B (zh) | 字幕编辑方法、装置和电子设备 | |
| CN108292301B (zh) | 上下文笔记记录 | |
| CN116137662B (zh) | 页面展示方法及装置、电子设备、存储介质和程序产品 | |
| WO2022089192A1 (zh) | 一种互动处理方法、装置、电子设备和存储介质 | |
| WO2021098571A1 (zh) | 基于在线文档评论的反馈方法、装置、设备及存储介质 | |
| WO2022017203A1 (zh) | 听写交互方法、装置和电子设备 | |
| US20240007718A1 (en) | Multimedia browsing method and apparatus, device and mediuim | |
| US10965743B2 (en) | Synchronized annotations in fixed digital documents | |
| CN104485115A (zh) | 发音评价设备、方法和系统 | |
| US11024199B1 (en) | Foreign language learning dictionary system | |
| WO2023134419A1 (zh) | 信息交互方法、装置、设备及存储介质 | |
| WO2023051294A9 (zh) | 道具处理方法、装置、设备及介质 | |
| CN111897976A (zh) | 一种虚拟形象合成方法、装置、电子设备及存储介质 | |
| US20240193206A1 (en) | Interaction method and apparatus, electronic device, and computer-readable storage medium | |
| CN113891168B (zh) | 字幕处理方法、装置、电子设备和存储介质 | |
| CN119003708B (zh) | 交互方法、装置及电子设备 | |
| CN114584716A (zh) | 图片处理方法、装置、设备及存储介质 | |
| CN114626332A (zh) | 内容展示方法、装置和电子设备 | |
| CN111260975B (zh) | 用于多媒体黑板教学互动的方法、装置、介质和电子设备 | |
| CN102956125B (zh) | 云端数码语音教学录音系统 | |
| CN111915174A (zh) | 基于电子绘本的小学生审辩性思维测评方法及系统 | |
| CN115269920A (zh) | 交互方法、装置、电子设备和存储介质 | |
| WO2022257777A1 (zh) | 多媒体处理方法、装置、设备及介质 | |
| CN112309390B (zh) | 信息交互方法和装置 |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 21846707 Country of ref document: EP Kind code of ref document: A1 |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 32PN | Ep: public notification in the ep bulletin as address of the adressee cannot be established |
Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205A DATED 020523) |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 21846707 Country of ref document: EP Kind code of ref document: A1 |