WO2024149183A1 - 文档显示方法、装置及电子设备 - Google Patents

文档显示方法、装置及电子设备 Download PDF

Info

Publication number
WO2024149183A1
WO2024149183A1 PCT/CN2024/071031 CN2024071031W WO2024149183A1 WO 2024149183 A1 WO2024149183 A1 WO 2024149183A1 CN 2024071031 W CN2024071031 W CN 2024071031W WO 2024149183 A1 WO2024149183 A1 WO 2024149183A1
Authority
WO
WIPO (PCT)
Prior art keywords
text
target
content
sub
document
Prior art date
Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
Ceased
Application number
PCT/CN2024/071031
Other languages
English (en)
French (fr)
Inventor
梁宵
Current Assignee (The listed assignees may be inaccurate. Google has not performed a legal analysis and makes no representation or warranty as to the accuracy of the list.)
Vivo Mobile Communication Co Ltd
Original Assignee
Vivo Mobile Communication Co Ltd
Priority date (The priority date is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the date listed.)
Filing date
Publication date
Application filed by Vivo Mobile Communication Co Ltd filed Critical Vivo Mobile Communication Co Ltd
Publication of WO2024149183A1 publication Critical patent/WO2024149183A1/zh
Anticipated expiration legal-status Critical
Ceased legal-status Critical Current

Links

Classifications

    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F3/00Input arrangements for transferring data to be processed into a form capable of being handled by the computer; Output arrangements for transferring data from processing unit to output unit, e.g. interface arrangements
    • G06F3/01Input arrangements or combined input and output arrangements for interaction between user and computer
    • G06F3/048Interaction techniques based on graphical user interfaces [GUI]
    • G06F3/0481Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance
    • G06F3/0483Interaction with page-structured environments, e.g. book metaphor
    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F3/00Input arrangements for transferring data to be processed into a form capable of being handled by the computer; Output arrangements for transferring data from processing unit to output unit, e.g. interface arrangements
    • G06F3/01Input arrangements or combined input and output arrangements for interaction between user and computer
    • G06F3/048Interaction techniques based on graphical user interfaces [GUI]
    • G06F3/0481Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance
    • G06F3/04817Interaction techniques based on graphical user interfaces [GUI] based on specific properties of the displayed interaction object or a metaphor-based environment, e.g. interaction with desktop elements like windows or icons, or assisted by a cursor's changing behaviour or appearance using icons
    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F3/00Input arrangements for transferring data to be processed into a form capable of being handled by the computer; Output arrangements for transferring data from processing unit to output unit, e.g. interface arrangements
    • G06F3/01Input arrangements or combined input and output arrangements for interaction between user and computer
    • G06F3/048Interaction techniques based on graphical user interfaces [GUI]
    • G06F3/0484Interaction techniques based on graphical user interfaces [GUI] for the control of specific functions or operations, e.g. selecting or manipulating an object, an image or a displayed text element, setting a parameter value or selecting a range
    • G06F3/04842Selection of displayed objects or displayed text elements
    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F3/00Input arrangements for transferring data to be processed into a form capable of being handled by the computer; Output arrangements for transferring data from processing unit to output unit, e.g. interface arrangements
    • G06F3/01Input arrangements or combined input and output arrangements for interaction between user and computer
    • G06F3/048Interaction techniques based on graphical user interfaces [GUI]
    • G06F3/0487Interaction techniques based on graphical user interfaces [GUI] using specific features provided by the input device, e.g. functions controlled by the rotation of a mouse with dual sensing arrangements, or of the nature of the input device, e.g. tap gestures based on pressure sensed by a digitiser
    • GPHYSICS
    • G06COMPUTING OR CALCULATING; COUNTING
    • G06FELECTRIC DIGITAL DATA PROCESSING
    • G06F40/00Handling natural language data
    • G06F40/20Natural language analysis
    • G06F40/279Recognition of textual entities
    • G06F40/289Phrasal analysis, e.g. finite state techniques or chunking
    • G06F40/295Named entity recognition

Definitions

  • the present application belongs to the field of communication technology, and specifically relates to a document display method, device and electronic equipment.
  • the purpose of the embodiments of the present application is to provide a document display method, device and electronic device, which can solve the problem of poor display effect of search results in related technologies.
  • an embodiment of the present application provides a document display method, comprising:
  • target content is displayed, where the target content is generated based on the writing progress of the document content and the N target objects;
  • the text progress is used to represent the characters and the distance between characters recorded in the document content. Writing order.
  • an embodiment of the present application provides a document display device, including:
  • a first receiving module is used to receive a first input from a user on N target objects in the document content displayed on the current interface, where N is a positive integer;
  • a first display module configured to display target content in response to the first input, wherein the target content is generated based on the writing progress of the document content and the N target objects;
  • the writing progress is used to represent the writing order of the characters and characters recorded in the document content.
  • an embodiment of the present application provides an electronic device, which includes a processor and a memory, wherein the memory stores programs or instructions that can be executed on the processor, and when the program or instructions are executed by the processor, the steps of the method described in the first aspect are implemented.
  • an embodiment of the present application provides a readable storage medium, on which a program or instruction is stored, and when the program or instruction is executed by a processor, the steps of the method described in the first aspect are implemented.
  • an embodiment of the present application provides a chip, comprising a processor and a communication interface, wherein the communication interface is coupled to the processor, and the processor is used to run a program or instruction to implement the method described in the first aspect.
  • an embodiment of the present application provides a computer program product, which is stored in a storage medium and is executed by at least one processor to implement the method described in the first aspect.
  • a first input of a user for N target objects in the document content displayed on the current interface is received, where N is a positive integer; in response to the first input, the target content is displayed, and the target content is generated based on the text progress of the document content and the N target objects; wherein the text progress is used to characterize the text order between the characters recorded in the document content.
  • the target content mainly uses the text progress as the time context to display the importance of the target object in the document content.
  • the user can know the importance of the target object in the document content, and because the target content uses the text progress as the time context to display The target object is displayed, so viewing the target content can also help users understand the development context of the document content, effectively improving the display effect of the target content.
  • FIG1 is a flow chart of a document display method provided in an embodiment of the present application.
  • FIG. 2 is a schematic diagram of the division of identification information provided in an embodiment of the present application.
  • FIG3 is one of the timing diagrams provided in an embodiment of the present application.
  • FIG4 is a second timing diagram provided in an embodiment of the present application.
  • FIG5 is a third timing diagram provided in an embodiment of the present application.
  • FIG6 is a fourth timing diagram provided in an embodiment of the present application.
  • FIG7 is a fifth timing diagram provided in an embodiment of the present application.
  • FIG8 is one of the graph relationship diagrams provided in the embodiments of the present application.
  • FIG. 9 is a second graph relationship diagram provided in an embodiment of the present application.
  • FIG10 is a third graph relationship diagram provided in an embodiment of the present application.
  • FIG11 is a structural diagram of a document display device provided in an embodiment of the present application.
  • FIG12 is a structural diagram of an electronic device provided in an embodiment of the present application.
  • FIG. 13 is a second structural diagram of the electronic device provided in an embodiment of the present application.
  • first”, “second”, etc. in the specification and claims of this application are used to distinguish similar objects, rather than to describe a specific order or sequence. It should be understood that the terms used in this way are interchangeable where appropriate, so that the embodiments of the present application can be implemented in an order other than those illustrated or described herein, and the objects distinguished by "first”, “second”, etc. are generally of the same type.
  • the number of objects is not limited, for example, the first object can be one or more.
  • “and/or” means at least one of the connected objects, and the character “/” generally means that the related objects are in an "or” relationship.
  • Entity It mainly refers to words or phrases in text data that are used to refer to things that exist in reality. For example, in the sentence “Xiao Ming is a student of A Middle School.”, “Xiao Ming”, “A Middle School”, and “Student” are all entities. Moreover, entities can also be classified. For example, “Xiao Ming” belongs to the “person” type, “A Middle School” belongs to the “school” type, and "student” belongs to the "occupation” type.
  • Event mainly refers to the events that occurred in the text. Generally speaking, a complete event includes: event, subject, object, time, place and other elements. For example, the event in “Xiao Ming attended class at Middle School A today” is “class”, the place where the event occurred is “Middle School A”, and the subject of the event is "Xiao Ming”.
  • Association relationship mainly refers to the association relationship between entities.
  • the association relationship can be a co-occurrence relationship or a sentiment relationship.
  • Co-occurrence relationship mainly refers to the co-occurrence of two entities in adjacent texts.
  • the co-occurrence relationship can be expressed in the form of intra-sentence co-occurrence, intra-paragraph co-occurrence, etc., and the co-occurrence relationship can also reflect the correlation between entities by the number of co-occurrences.
  • Sentiment analysis mainly through a text analysis technology to analyze the emotional relationship contained in the text, and even set the emotional level for the emotional relationship. For example, "How can you be so heartless?! contains a negative emotion.
  • emotion levels include but are not limited to “like”, “close”, “neutral”, “distant” and “dislike”.
  • Figure 1 is a flow chart of the document display method provided in the embodiment of the present application.
  • the document display method provided in the embodiment of the present application can be applied to mobile terminals such as mobile phones and tablet computers, and the document display method includes the following steps:
  • Step 101 Receive a first input from a user on N target objects in the document content displayed on the current interface.
  • the document content may be an electronic book such as an online novel or a biography, or may be a PDF file or a Word file. Moreover, the document content in this embodiment may refer to the content of the file or book currently being viewed by the user.
  • the target object may be a word or phrase included in the portion of the document content displayed on the current interface. Moreover, in this embodiment, the target object may refer to the aforementioned entity.
  • the first input is mainly used to trigger a subsequent process, which may be a long press operation or a double-click operation on the target object.
  • the first input may also be other triggering operations, which are not specifically limited here.
  • N is a positive integer, that is, in the present application, it is possible to receive only the first input for one target object, or to receive the first input for two or more target objects.
  • Step 102 In response to the first input, display target content, where the target content is generated based on the writing progress of the document content and the N target objects.
  • the text progress can be used to represent the text order of the characters and characters recorded in the document content.
  • the order of the first character and the last character recorded in the document content can be represented as the progress of the text.
  • the document content includes the first chapter, the second chapter and the third chapter
  • the first chapter includes pages 1-6 of the document content
  • the second chapter includes pages 7-11 of the document content
  • the third chapter includes pages 12-18 of the document content.
  • the page number sequence of pages 1-18 of the document content can be understood as the order of the text of the document content and can be represented as the progress of the text.
  • the writing progress can be based not only on pages, but also on the separators of document content such as chapters, paragraphs, sentences, etc.
  • the current reading progress of the document content can also be used as the writing progress, that is, the content included from the first character of the document content to the content displayed on the current interface can be used as the writing progress of the document content, so as to generate target content based on the current reading progress and target object.
  • the target content may be a timeline showing the frequency of occurrence of the target object in the document content using the writing progress as a time axis.
  • the target content may be in the form of a table or a coordinate graph.
  • the target content can display the importance of the target object in the document content with the progress of the text as the time context. In this way, by viewing the displayed target content, the user can know the importance of the target object in the document content, and because the target content displays the target object with the progress of the text as the time context, viewing the target content can also help the user understand the development context of the document content, effectively improving the display effect of the target content.
  • the target content can be displayed on the current interface in the form of a floating window.
  • the document content can be divided into chapters, pages, paragraphs, and sentences.
  • chapter information can be obtained based on the keywords of the document content (such as "Chapter 3 XXX"); paragraph information can be obtained based on the line breaks of the document content; page number information can be obtained based on the page number of the document content; sentence information can be obtained based on the punctuation marks of the document content (".!? --, etc.).
  • the chapter number, page number, paragraph number, and sentence number of each sentence obtained can all be increased sequentially starting from 1.
  • the numbering information of the chapter, page number, paragraph, sentence, etc. of the document content can be used as the identification information of the sentence, and the identification information of each sentence is unique.
  • the identification information is “Chapter 1, Page 15, Paragraph 78, Sentence 235”, and the corresponding sentence in the document content is “What are you doing here?”
  • the identification information is “Chapter 1, Page 15, Paragraph 80, Sentence 237”, and the corresponding sentence in the document content is “Xiao Ming has returned home from school.”.
  • chapters, page numbers, paragraphs, and sentences all start increasing from 1 and do not restart; page numbers, paragraphs, and sentence numbers are not restarted due to chapter numbers; paragraphs are not restarted due to changes in chapters or page numbers; sentences are not restarted due to changes in chapters, page numbers, or paragraphs.
  • named entity recognition can be performed for each sentence, including name recognition, location recognition, event recognition, organization recognition, and proper name recognition.
  • the recognition results include: entity name, entity position in the sentence, entity type, etc.
  • the named entities can be identified to obtain: entity type-person position-1, entity type-person position-4, entity type-person position 7.
  • the identified named entities can be stored in the database module 14, for example, in the form of Table 1.
  • the number of occurrences of entities such as people, places, and locations in the document content can be counted, and can be sorted in reverse order and stored in the database.
  • the counted entity information can be stored in the form of Table 2.
  • the number of entity occurrences of characters, places, events, etc. in each chapter and page can be counted according to the identification information of the sentence and stored in the database.
  • the number of entity occurrences in each chapter can be shown in Table 3, and the number of entity occurrences in each page can be shown in Table 4.
  • the document content can also be analyzed in time series.
  • the document content can be divided into multiple time windows, and then sentiment analysis is performed on each time window, and the corresponding sentiment analysis results are stored until all time windows complete the sentiment analysis.
  • the document content can be divided into several time windows.
  • the nth paragraph corresponds to the nth window.
  • the time window length corresponding to the first window and the last window is 2 (that is, when X is 3), because they are missing the previous paragraph and the next paragraph respectively.
  • sentiment has direction. A's sentiment towards B and B's sentiment towards A are different. Must be consistent.
  • entity sentiment analysis based on deep learning can be used to classify emotions into multiple levels. For example, emotions can be divided into 5 levels: disgust (-2), separation (-1), neutral (0), closeness (1), and liking (2), and the deep learning model can be used to determine which level the mutual emotions between two entities belong to.
  • the sentiment analysis results obtained after sentiment analysis can be stored in the database in the form of Table 5.
  • the above process can be executed on a client or on a server such as a cloud platform.
  • the client may send a request message to the server, and the request message includes the current page number information.
  • the server may return the identified entities such as people, places, events, etc. to the client based on the received request message, and mark them with identifiers (color block identifiers, etc.) in the reader (which may include but is not limited to mobile phones), and then count the full text of the entity based on the statistics, and color the entity identifier according to the count (for example, rank in descending order of count, divide the entities in the text into 5 equal parts, the one with the highest count is colored the darkest, and the one with the lowest count is colored the lightest).
  • identifiers color block identifiers, etc.
  • the target object in the present application may be the entity mentioned in the above-mentioned implementation mode, that is, the target object may be entity information such as a person, a place, an event, etc. in the document content.
  • N is greater than 1; and the display target content includes:
  • the related content includes at least one of the number of co-occurrences and the emotion level.
  • the preset separator can be a chapter separator, page number separator, paragraph separator or sentence separator of the document content.
  • the user can reasonably set the preset separator based on actual needs (such as the length of the document content) to reduce the amount of calculation required to divide the sub-text.
  • the division of the M subtexts can be performed on the client side or on the server side.
  • the association relationship between the at least two target objects can be obtained based on the writing progress, so that the user can better understand the development context of the document content based on the association relationship between the at least two target objects, thereby improving the user's reading efficiency.
  • the development context of the document content can be better understood based on the associated content such as the number of co-occurrences and the sentiment level of the two target objects.
  • N 2
  • the association relationship between the two target objects in any time window can be obtained.
  • acquiring the association relationship between the N target objects in any one of the M sub-texts based on the text progress includes:
  • a first association relationship between the N target objects in the first sub-text is obtained; wherein, when the first association relationship indicates that there is a mutual association between the N target objects, the first association relationship is determined as the association relationship between the N target objects in the first sub-text; when the first association relationship indicates that there is no mutual association between the N target objects, a second association relationship between the N target objects in the second sub-text is obtained, and the second association relationship is determined as the association relationship between the N target objects in the first sub-text; the first sub-text and the second sub-text are adjacent sub-texts in the M sub-texts with the writing progress as the time axis, and the writing order of the first sub-text is located after the second sub-text.
  • the association relationship among the N target objects in the first subtext may be determined based on the first association relationship among the N target objects in the first subtext.
  • the first association relationship when the first association relationship represents that there is a mutual association between N target objects, the first association relationship can be determined as the association relationship of the N target objects in the first sub-text; and when the first association relationship represents that there is no mutual association between the N target objects, the second association relationship of the N target objects in the second sub-text can be determined as the association relationship of the N target objects in the first sub-text, so that the association relationship of the N target objects in the document content has continuity and the purpose of improving the user's reading experience is achieved.
  • the association relationship between the N target objects in the first subtext can be determined as the association relationship of the N target objects in the first subtext (first text).
  • a time series analysis diagram of the emotional relationship of the target object can be displayed on the display interface of the mobile terminal.
  • the target content is a time sequence diagram of the target object in the document content
  • the timing diagram includes a horizontal axis, a vertical axis and a display curve
  • the horizontal axis is based on the writing progress
  • the vertical axis is based on the number of occurrences of the target object in any writing unit
  • the display curve is a curve formed by connecting the number of occurrences of the target object in each writing unit
  • the writing unit is obtained by dividing the document content based on the writing progress.
  • the timing diagram may display the occurrence frequency of a single target object in the document content based on the text progress, and the timing diagram may also display the association relationship between at least two target objects in the document content based on the text progress.
  • the timing diagram further includes a pull-bar icon displayed on the horizontal axis, and the position of the pull-bar icon on the horizontal axis corresponds to the position of the text unit displayed on the current interface on the text progress;
  • the method further includes:
  • the user can quickly jump to the text paragraph including the target object in the target text unit corresponding to the updated pull-bar icon, so as to improve the user's reading efficiency.
  • the second input may be a long press operation or a double click operation on the pull bar icon.
  • the mobile phone when a user is reading document content through a mobile phone, when the user receives a click operation or a long press operation on an entity word (target object) such as a person, place, event, etc. displayed on the display screen of the mobile phone, the mobile phone will send a request message to the server through the network, and the request message includes the entity name on which the above-mentioned click operation or long press operation acts; after receiving the request message, the server will respond to the request message and obtain the count information of the entity in the page number and chapter from the database module of the server, and the server will send the obtained count information to the mobile phone so that the mobile phone can generate the target content based on the received count information.
  • an entity word target object
  • the server after receiving the request message, the server will respond to the request message and obtain the count information of the entity in the page number and chapter from the database module of the server, and the server will send the obtained count information to the mobile phone so that the mobile phone can generate the target content based on the received count information.
  • the target content can be displayed in the form of a floating window above the content displayed on the current interface to show the frequency of occurrence of the target object (such as entities such as people, places, events, etc.) in each page and chapter in the full text of the document content.
  • the target content can be a frequency time series diagram, and includes two coordinate axes and a display curve. The horizontal axis of the two coordinate axes is used to indicate the page number, and the chapter is marked on the starting page of the chapter; the vertical axis is used to indicate the count of the target object of the current page.
  • the position of the current page in the full text can be displayed in the frequency time series diagram in the form of a pull rod icon (such as a pull rod with a straight line and a round ball). And by pulling the pull rod icon, it is also possible to adjust the reading position and jump directly to the page number indicated by the pull rod icon, that is, it is possible to jump to the text paragraph including the target object in the target text unit corresponding to the updated pull rod icon.
  • a pull rod icon such as a pull rod with a straight line and a round ball
  • the target content may also be a time sequence diagram of the association relationship between at least two target objects.
  • the sentiment time sequence diagram uses pages as the horizontal axis (i.e., the order of the text of the document content is the horizontal axis), and the chapters are marked on the starting page of the chapters, and the sentiment scores of at least two target objects on each page are used as the vertical axis.
  • the current page will be displayed in the emotional time sequence diagram in the form of a lever icon (such as a straight line plus a ball lever).
  • the reading position can be adjusted and the page number indicated by the lever icon can be directly jumped to the text paragraph including the target object in the target text unit corresponding to the updated lever icon.
  • the target content when the first input acts on at least two target objects, the target content may include not only the emotion timing diagrams of the at least two target objects, but also the frequency timing diagrams of the at least two target objects, as shown in FIG. 7 for details.
  • the timing diagram further includes a map icon of the target object
  • the method further includes:
  • the associated object is an object in the document content that has an association relationship with the target object, and the degree of association between the target object and the associated object is negatively correlated with the connection length, and the connection length is the distance length between the target object and the associated object in the graph relationship diagram.
  • a graph icon of the target object can also be displayed in the timing diagram.
  • a graph relationship diagram of the target object and its associated objects can also be displayed, so that users can better understand the development context of the document content and achieve efficient reading of the document content.
  • the third input can be a long press operation or a double click operation on the map icon.
  • the mobile terminal when a click operation is received on the map icon of "View Complete Character Map" shown in Figure 4, the mobile terminal will send a request message to the server, where the request message is the target object shown in Figure 4 (and marked as "central entity"), to obtain the underlying map data of the target object up to the current page.
  • the underlying data of the graph includes: the acquired entity recognition data, and it is possible to request to obtain all entities that co-occur with the central entity within a segment; and the acquired sentiment relationship data, and it is possible to request to obtain the relationship data between all entities that co-occur with the central entity within a segment and the central entity.
  • the graph data of the current page can be calculated based on the underlying data of the graph, thereby generating a graph relationship diagram. It can be divided into two steps. The first is to calculate all entities and their co-occurrence times that co-occur with the central entity up to the current page. The second is to calculate the relationship between the central entity and its co-occurring entities up to the current page according to the interpolation method mentioned in the above content.
  • the graph relationship diagram in this embodiment can be shown in FIG8.
  • the relationship between the target object (central entity) "Xiao Ming" and other entities can be divided into three levels: strong, medium, and weak in the graph, and the positions of all other entities can refer to the method shown in FIG8.
  • the sentiment levels between entities can also be marked with lines so that users can view the overall context of the document content.
  • the central entity e0 can be placed at the center of the graph relationship diagram, and the long side length of the graph relationship diagram can be set to L;
  • the strong relationship area can be set to a circle with the central entity as the center and L/6 as the radius;
  • the medium relationship area can be set to a circle with the central entity as the center and L/3 as the radius;
  • the weak relationship area can be set to the area outside the strong relationship area and the medium relationship area.
  • the example of ex and the central entity can be limited to Lx, and ex can be evenly placed in the graph relationship diagram.
  • the central entity can be set as the clicked entity.
  • the graph relationship diagram with “Xiao Ming” as the central entity shown in FIG8 can be switched to the graph relationship diagram with “Xiao Wang” as the central entity as shown in FIG9 , and the complete character graph of “Xiao Wang” can be obtained.
  • a lever icon is also displayed. By pulling or clicking the lever icon, the graph relationship diagram can be adjusted. For example, pull the lever icon in Figure 9 to Chapter 6, and the updated graph relationship diagram is shown in Figure 10.
  • “Xiao Wang” and “Xiao Li” are also familiar with each other, and the emotional relationship has become close to each other; but “Xiao Wang” is alienated from “Xiao Zhao”, and the emotional relationship has become repulsive.
  • the mobile terminal can be not only a terminal device such as a mobile phone or a tablet, but also a VR device, that is, the above solution can also be implemented in a VR scenario.
  • the document content may be divided into fields or categories to improve the recognition and analysis process of the document content.
  • online encyclopedias of some entities and other content can also be combined to assist users in reading, so as to help users better understand the development context of the document content and achieve efficient reading.
  • the document display method of the embodiment of the present application receives a first input from a user for N target objects in the document content displayed on the current interface, where N is a positive integer; in response to the first input, the target content is displayed, and the target content is generated based on the text progress of the document content and the N target objects; wherein the text progress is used to characterize the text order between the characters recorded in the document content.
  • N is a positive integer
  • the text progress is used to characterize the text order between the characters recorded in the document content.
  • the document display method provided in the embodiment of the present application can be executed by a document display device.
  • the document display device provided in the embodiment of the present application is described by taking the document display method executed by the document display device as an example.
  • FIG. 11 is a structural diagram of a document display device provided in an embodiment of the present application. As shown in FIG. 11 , the device 1100 includes:
  • the first receiving module 1101 is used to receive a first input from a user regarding N target objects in the document content displayed on the current interface, where N is a positive integer;
  • a first display module 1102 is configured to display target content in response to the first input, where the target content is generated based on the writing progress of the document content and the N target objects;
  • the writing progress is used to represent the writing order of the characters and characters recorded in the document content.
  • N is greater than 1; the first display module 1102 includes:
  • An acquisition unit configured to acquire, based on the text progress, an association relationship between the N target objects in any one of the M subtexts, wherein the M subtexts are obtained by dividing the document content based on a preset delimiter, and M is an integer greater than 1;
  • a display unit used to display the associated content of the N target objects in each subtext based on the text progress as a time axis, wherein the associated content is determined based on the association relationship between the N target objects in each subtext;
  • the related content includes at least one of the number of co-occurrences and the emotion level.
  • the acquisition unit is specifically used to acquire a first association relationship between the N target objects in the first sub-text when the first sub-text is not the first text of the document content; wherein, when the first association relationship represents that there is a mutual association between the N target objects, the first association relationship is determined as the association relationship between the N target objects in the first sub-text; when the first association relationship represents that there is no mutual association between the N target objects, a second association relationship between the N target objects in the second sub-text is acquired, and the second association relationship is determined as the association relationship between the N target objects in the first sub-text; the first sub-text and the second sub-text are adjacent sub-texts among the M sub-texts with the writing progress as the timeline, and the writing order of the first sub-text is located after the second sub-text.
  • the target content is a time sequence diagram of the target object in the document content
  • the timing diagram includes a horizontal axis, a vertical axis and a display curve
  • the horizontal axis is based on the writing progress
  • the vertical axis is based on the number of occurrences of the target object in any writing unit
  • the display curve is a curve formed by connecting the number of occurrences of the target object in each writing unit
  • the writing unit is obtained by dividing the document content based on the writing progress.
  • the timing diagram further includes a pull bar icon displayed on the horizontal axis, and the pull bar icon The position of the bar icon on the horizontal axis corresponds to the position of the text unit displayed on the current interface on the text progress;
  • the device 1100 further includes:
  • a second receiving module used for receiving a second input to the pull bar icon
  • a determination module configured to determine a target text unit corresponding to the pull bar icon in response to the second input
  • a switching display module is used to switch to and display the text paragraph including the target object in the target text unit.
  • the timing diagram further includes a map icon of the target object
  • the device 1100 further includes:
  • a third receiving module used for receiving a third input of the atlas icon
  • a second display module configured to display a graph relationship between the target object and associated objects in response to the third input
  • the associated object is an object in the document content that has an associated relationship with the target object, and the degree of association between the target object and the associated object is negatively correlated with the connection length, and the connection length is the distance length between the target object and the associated object in the graph relationship.
  • the document display device in the embodiment of the present application can be an electronic device or a component in the electronic device, such as an integrated circuit or a chip.
  • the electronic device can be a terminal or other devices other than a terminal.
  • the electronic device can be a mobile phone, a tablet computer, a laptop computer, a PDA, a vehicle-mounted electronic device, a mobile Internet device (Mobile Internet Device, MID), an augmented reality (Augmented Reality, AR)/virtual reality (Virtual Reality, VR) device, a robot, a wearable device, an ultra-mobile personal computer (Ultra-Mobile Personal Computer, UMPC), a netbook or a personal digital assistant (Personal Digital Assistant, PDA), etc.
  • Network Attached Storage Network Attached Storage
  • PC Personal Computer
  • TV Television
  • teller machine a self-service machine, etc., which is not specifically limited in the embodiment of the present application.
  • the document display device in the embodiment of the present application may be a device having an operating system.
  • the system may be an Android operating system, an iOS operating system, or other possible operating systems, which are not specifically limited in the embodiments of the present application.
  • the document display device provided in the embodiment of the present application can implement each process implemented by the method embodiment of Figure 1, and will not be described again here to avoid repetition.
  • an embodiment of the present application also provides an electronic device 1200, including a processor 1201 and a memory 1202, and the memory 1202 stores a program or instruction that can be executed on the processor 1201.
  • the program or instruction is executed by the processor 1201
  • the various steps of the above-mentioned document display method embodiment are implemented, and the same technical effect can be achieved. To avoid repetition, it will not be repeated here.
  • the electronic devices in the embodiments of the present application include the mobile electronic devices and non-mobile electronic devices mentioned above.
  • the electronic device 1300 includes but is not limited to: a radio frequency unit 1301, a network module 1302, an audio output unit 1303, an input unit 1304, a sensor 1305, a display unit 1306, a user input unit 1307, an interface unit 1308, a memory 1309, and a processor 1310 and other components.
  • the electronic device 1300 may also include a power source (such as a battery) for supplying power to each component, and the power source may be logically connected to the processor 1310 through a power management system, so that the power management system can manage charging, discharging, and power consumption management.
  • a power source such as a battery
  • the electronic device structure shown in FIG13 does not constitute a limitation on the electronic device, and the electronic device may include more or fewer components than shown, or combine certain components, or arrange components differently, which will not be described in detail here.
  • the user input unit 1307 is used to receive a first input from a user on N target objects in the document content displayed on the current interface, where N is a positive integer;
  • a display unit 1306 is configured to display target content in response to the first input, where the target content is generated based on the writing progress of the document content and the N target objects;
  • the writing progress is used to represent the writing order of the characters and characters recorded in the document content.
  • the associated relationship between the N target objects in any sub-text of the document, the M sub-texts are obtained by dividing the document content based on a preset delimiter, and M is an integer greater than 1;
  • the display unit 1306 is used to display the associated content of the N target objects in each sub-text with the writing progress as the time axis, and the associated content is determined based on the associated relationship between the N target objects in each sub-text; wherein the associated content includes at least one of the number of co-occurrences and the sentiment level.
  • processor 1310 is used to obtain a first association relationship between the N target objects in the first sub-text when the first sub-text is not the first text of the document content; wherein, when the first association relationship represents that there is a mutual association between the N target objects, the first association relationship is determined as the association relationship between the N target objects in the first sub-text; when the first association relationship represents that there is no mutual association between the N target objects, a second association relationship between the N target objects in the second sub-text is obtained, and the second association relationship is determined as the association relationship between the N target objects in the first sub-text; the first sub-text and the second sub-text are adjacent sub-texts among the M sub-texts with the writing progress as the timeline, and the writing order of the first sub-text is located after the second sub-text.
  • the target content is a time sequence diagram of the target object in the document content
  • the timing diagram includes a horizontal axis, a vertical axis and a display curve
  • the horizontal axis is based on the writing progress
  • the vertical axis is based on the number of occurrences of the target object in any writing unit
  • the display curve is a curve formed by connecting the number of occurrences of the target object in each writing unit
  • the writing unit is obtained by dividing the document content based on the writing progress.
  • the timing diagram also includes a pull bar icon displayed on the horizontal axis, and the position of the pull bar icon on the horizontal axis corresponds to the position of the text unit displayed on the current interface on the text progress; a user input unit 1307 is used to receive a second input to the pull bar icon; a processor 1310 is used to determine a target text unit corresponding to the pull bar icon in response to the second input; and a display unit 1306 is used to switch to and display the text paragraph including the target object in the target text unit.
  • the sequence diagram further includes a map icon of the target object; the user input unit 1307, Used to receive a third input to the graph icon; display unit 1306, used to display the graph relationship between the target object and the associated object in response to the third input; wherein the associated object is an object in the document content that has an associated relationship with the target object, and the degree of association between the target object and the associated object is negatively correlated with the length of the connection line, and the length of the connection line is the distance length between the target object and the associated object in the graph relationship.
  • the associated object is an object in the document content that has an associated relationship with the target object, and the degree of association between the target object and the associated object is negatively correlated with the length of the connection line, and the length of the connection line is the distance length between the target object and the associated object in the graph relationship.
  • the input unit 1304 may include a graphics processing unit (GPU) 13041 and a microphone 13042, and the graphics processor 13041 processes the image data of the static picture or video obtained by the image capture device (such as a camera) in the video capture mode or the image capture mode.
  • the display unit 1306 may include a display panel 13061, and the display panel 13061 may be configured in the form of a liquid crystal display, an organic light emitting diode, etc.
  • the user input unit 1307 includes a touch panel 13071 and at least one of other input devices 13072.
  • the touch panel 13071 is also called a touch screen.
  • the touch panel 13071 may include two parts: a touch detection device and a touch controller.
  • Other input devices 13072 may include, but are not limited to, a physical keyboard, function keys (such as a volume control key, a switch key, etc.), a trackball, a mouse, and a joystick, which will not be repeated here.
  • the memory 1309 can be used to store software programs and various data.
  • the memory 1309 may mainly include a first storage area for storing programs or instructions and a second storage area for storing data, wherein the first storage area may store an operating system, an application program or instructions required for at least one function (such as a sound playback function, an image playback function, etc.), etc.
  • the memory 1309 may include a volatile memory or a non-volatile memory, or the memory 1309 may include both volatile and non-volatile memories.
  • the non-volatile memory may be a read-only memory (ROM), a programmable read-only memory (PROM), an erasable programmable read-only memory (EPROM), an electrically erasable programmable read-only memory (EEPROM), or a flash memory.
  • Volatile memory can be random access memory (RAM), static random access memory (SRAM), dynamic random access memory (DRAM), synchronous dynamic random access memory (SDRAM), double data rate synchronous dynamic random access memory (Double Data Rate SDRAM, DDRSDRAM), Enhanced SDRAM (ESDRAM), Synch link DRAM (SLDRAM) and Direct Rambus RAM (DRRAM).
  • RAM random access memory
  • SRAM static random access memory
  • DRAM dynamic random access memory
  • SDRAM synchronous dynamic random access memory
  • DDRSDRAM synchronous dynamic random access memory
  • ESDRAM Enhanced SDRAM
  • SLDRAM Synch link DRAM
  • DRRAM Direct Rambus RAM
  • the memory 1309 in the embodiment of the present application includes but is not limited to these
  • the processor 1310 may include one or more processing units; optionally, the processor 1310 integrates an application processor and a modem processor, wherein the application processor mainly processes operations related to an operating system, a user interface, and application programs, and the modem processor mainly processes wireless communication signals, such as a baseband processor. It is understandable that the modem processor may not be integrated into the processor 1310.
  • An embodiment of the present application also provides a readable storage medium, which may be volatile or non-volatile.
  • a program or instruction is stored on the readable storage medium.
  • the program or instruction is executed by a processor, the various processes of the above-mentioned document display method embodiment are implemented and the same technical effect can be achieved. To avoid repetition, it will not be repeated here.
  • the processor is the processor in the electronic device described in the above embodiment.
  • the readable storage medium includes a computer readable storage medium, such as a computer read-only memory ROM, a random access memory RAM, a magnetic disk or an optical disk.
  • An embodiment of the present application further provides a chip, which includes a processor and a communication interface, wherein the communication interface is coupled to the processor, and the processor is used to run programs or instructions to implement the various processes of the above-mentioned document display method embodiment, and can achieve the same technical effect. To avoid repetition, it will not be repeated here.
  • the chip mentioned in the embodiments of the present application can also be called a system-level chip, a system chip, a chip system or a system-on-chip chip, etc.
  • An embodiment of the present application provides a computer program product, which is stored in a storage medium.
  • the program product is executed by at least one processor to implement the various processes of the document display method embodiment described above, and can achieve the same technical effect. To avoid repetition, it will not be repeated here.
  • the technical solution of the present application can be embodied in the form of a computer product, which is stored in a storage medium (such as ROM/RAM, a disk, or an optical disk), and includes a number of instructions for a terminal (which can be a mobile phone, a computer, a server, or a network device, etc.) to execute the methods described in each embodiment of the present application.
  • a storage medium such as ROM/RAM, a disk, or an optical disk

Landscapes

  • Engineering & Computer Science (AREA)
  • Theoretical Computer Science (AREA)
  • General Engineering & Computer Science (AREA)
  • Physics & Mathematics (AREA)
  • General Physics & Mathematics (AREA)
  • Human Computer Interaction (AREA)
  • Health & Medical Sciences (AREA)
  • Artificial Intelligence (AREA)
  • Audiology, Speech & Language Pathology (AREA)
  • Computational Linguistics (AREA)
  • General Health & Medical Sciences (AREA)
  • User Interface Of Digital Computer (AREA)
  • Digital Computer Display Output (AREA)

Abstract

本申请公开了一种文档显示方法、装置及电子设备,属于通信技术领域。该文档显示方法包括:接收用户对当前界面显示的文档内容中的N个目标对象的第一输入,N为正整数;响应于所述第一输入,显示目标内容,所述目标内容是基于所述文档内容的行文进度和所述N个目标对象生成得到的;其中,所述行文进度用于表征所述文档内容所记载的字符和字符之间的行文顺序。

Description

文档显示方法、装置及电子设备
相关申请的交叉引用
本申请主张在2023年1月13日在中国提交的中国专利申请No.202310068177.4的优先权,其全部内容通过引用包含于此。
技术领域
本申请属于通信技术领域,具体涉及一种文档显示方法、装置及电子设备。
背景技术
随着技术的发展,手机、平板等电子设备逐渐成为用户阅读文档的新媒介。用户通过借助于这些电子设备,可以获得海量的专业著作、小说等文档内容。然而,当用户需要获取文档内容中的特定人物或特定事件的时候,通常是通过文内搜索的方式进行获取。但当特定人物或者特定事件在文档内容中出现的次数较多时,难以将获取到的搜索结果直观的展示给用户。
可见,相关技术中的搜索结果存在展示效果差的问题。
发明内容
本申请实施例的目的是提供一种文档显示方法、装置及电子设备,能够解决相关技术中的搜索结果存在的展示效果差的问题。
为了解决上述技术问题,本申请是这样实现的:
第一方面,本申请实施例提供了一种文档显示方法,包括:
接收用户对当前界面显示的文档内容中的N个目标对象的第一输入,N为正整数;
响应于所述第一输入,显示目标内容,所述目标内容是基于所述文档内容的行文进度和所述N个目标对象生成得到的;
其中,所述行文进度用于表征所述文档内容所记载的字符和字符之间的 行文顺序。
第二方面,本申请实施例提供了一种文档显示装置,包括:
第一接收模块,用于接收用户对当前界面显示的文档内容中的N个目标对象的第一输入,N为正整数;
第一显示模块,用于响应于所述第一输入,显示目标内容,所述目标内容是基于所述文档内容的行文进度和所述N个目标对象生成得到的;
其中,所述行文进度用于表征所述文档内容所记载的字符和字符之间的行文顺序。
第三方面,本申请实施例提供了一种电子设备,该电子设备包括处理器和存储器,所述存储器上存储有可在所述处理器上运行的程序或指令,所述程序或指令被所述处理器执行时实现如第一方面所述的方法的步骤。
第四方面,本申请实施例提供了一种可读存储介质,所述可读存储介质上存储程序或指令,所述程序或指令被处理器执行时实现如第一方面所述的方法的步骤。
第五方面,本申请实施例提供了一种芯片,所述芯片包括处理器和通信接口,所述通信接口和所述处理器耦合,所述处理器用于运行程序或指令,实现如第一方面所述的方法。
第六方面,本申请实施例提供一种计算机程序产品,该程序产品被存储在存储介质中,该程序产品被至少一个处理器执行以实现如第一方面所述的方法。
在本申请实施例中,通过接收用户对当前界面显示的文档内容中的N个目标对象的第一输入,N为正整数;响应于所述第一输入,显示目标内容,所述目标内容是基于所述文档内容的行文进度和所述N个目标对象生成得到的;其中,所述行文进度用于表征所述文档内容所记载的字符和字符之间的行文顺序。目标内容主要以行文进度为时间脉络,对目标对象在文档内容中的重要性进行展示。这样用户通过查看显示的目标内容,就可以知晓目标对象在文档内容中的重要性,而且由于目标内容是以行文进度为时间脉络来展 示目标对象的,因此通过查看目标内容还可以帮助用户理解文档内容的发展脉络,有效的提升了目标内容的展示效果。
附图说明
图1是本申请实施例提供的文档显示方法的流程图;
图2是本申请实施例提供的标识信息的划分示意图;
图3是本申请实施例提供的时序图之一;
图4是本申请实施例提供的时序图之二;
图5是本申请实施例提供的时序图之三;
图6是本申请实施例提供的时序图之四;
图7是本申请实施例提供的时序图之五;
图8是本申请实施例提供的图谱关系图的之一;
图9是本申请实施例提供的图谱关系图的之二;
图10是本申请实施例提供的图谱关系图的之三;
图11是本申请实施例提供的文档显示装置的结构图;
图12是本申请实施例提供的电子设备的结构图之一;
图13是本申请实施例提供的电子设备的结构图之二。
具体实施方式
下面将结合本申请实施例中的附图,对本申请实施例中的技术方案进行清楚、完整地描述,显然,所描述的实施例是本申请一部分实施例,而不是全部的实施例。基于本申请中的实施例,本领域普通技术人员在没有作出创造性劳动前提下所获得的所有其他实施例,都属于本申请保护的范围。
本申请的说明书和权利要求书中的术语“第一”、“第二”等是用于区别类似的对象,而不用于描述特定的顺序或先后次序。应该理解这样使用的术语在适当情况下可以互换,以便本申请的实施例能够以除了在这里图示或描述的那些以外的顺序实施,且“第一”、“第二”等所区分的对象通常为一类, 并不限定对象的个数,例如第一对象可以是一个,也可以是多个。此外,说明书以及权利要求中“和/或”表示所连接对象的至少其中之一,字符“/”,一般表示前后关联对象是一种“或”的关系。
下面先对本申请中涉及的一些关键词进行说明。
实体:主要是指文本数据中,用于指代现实中存在的事物的词汇或者短语。例如“小明是A中学的学生。”这句话中,“小明”、“A中学”、“学生”都是实体。而且,实体还可以进行分类,例如,“小明”属于“人物”类型,“A中学”属于“学校”类型,“学生”属于“职业”类型。
事件:主要是指文本中发生的事件。一般来说,一个完整的事件包含:事件、主体、客体、时间、地点等要素。例如:“小明今天在A中学上课”里面的事件就是“上课”,事件发生的地点是“A中学”,事件的主体是“小明”。
关联关系:主要是指实体之间的关联关系。在本申请中,关联关系可以共现关系,也可以是情感关系。
共现关系:主要是指两个实体在临近的文本中同时出现。在本申请中,共现关系可以表现为句内共现、段内共现等形式,且共现关系还可以共现次数反映出实体之间的相关性。
情感分析:主要是通过一种文本分析技术,以分析出文本包含的情感关系,甚至可以给情感关系设置情感级别。例如,“你这人怎地恁地无情?!”包含的是一种负面情感。
可以理解的是,情感级别包括但不限于“喜欢”、“亲近”、“中性”、“疏离”和“厌恶”等。
下面结合附图,通过具体的实施例及其应用场景对本申请实施例提供的文档显示方法进行详细地说明。
请参见图1,图1是本申请实施例提供的文档显示方法的流程图。如图1所示,本申请实施例提供的文档显示方法可以应用在手机、平板电脑等移动终端,且该文档显示方法包括以下步骤:
步骤101、接收用户对当前界面显示的文档内容中的N个目标对象的第一输入。
本实施方式中,文档内容可以是网络小说、人物传记等电子书,也可以是PDF文件或者Word文件。而且,本实施方式中的文档内容可以是指用户当前正在查阅的文件或书籍等内容。
目标对象可以是文档内容在当前界面显示的部分所包括的词汇或者短语。而且,在本实施方式中,目标对象可以是指前述提到的实体。
另外,第一输入主要用于触发后续流程,其可以是作用在目标对象上的长按操作或者双击操作等。
一些实施方式中,第一输入还可以是其他触发性操作,在此不做具体限定。
可以理解的是,N为正整数,即在本申请中,既可以仅接收针对一个目标对象的第一输入,也可以接收针对两个或者多个目标对象的第一输入。
步骤102、响应于所述第一输入,显示目标内容,所述目标内容是基于所述文档内容的行文进度和所述N个目标对象生成得到的。
本实施方式中,行文进度可以用于表征文档内容所记载的字符和字符之间的行文顺序。
示例性的,文档内容所记载的第一个字符和最后一个字符之间的行文顺序可以表征为行文进度。比如,文档内容包括第一章节、第二章节和第三章节,第一章节包括文档内容的第1-6页,第二章节包括文档内容的第7-11页,第三章节包括文档内容的第12-18页,则文档内容的第1-18页的页码顺序可以理解成该文档内容的行文顺序,并可以表征为行文进度。
可以理解的是,行文进度不仅可以以页为时间脉络,还可以以章、段、句等文档内容的分隔符为时间脉络。
在另一些实施方式中,还可以将文档内容当前的阅读进度作为行文进度,即可以将文档内容的第一字符至当前界面显示的内容之间所囊括的内容作为文档内容的行文进度,以便基于当前的阅读进度和目标对象生成目标内容。
其中,目标内容可以是以行文进度作为时间轴来展示目标对象在文档内容中的出现频次。而且,目标内容可以是表格的形式,也可以是坐标图的形式。
本实施方式中,目标内容可以以行文进度为时间脉络,对目标对象在文档内容中的重要性进行展示。这样用户通过查看显示的目标内容,就可以知晓目标对象在文档内容中的重要性,而且由于目标内容是以行文进度为时间脉络来展示目标对象的,因此通过查看目标内容还可以帮助用户理解文档内容的发展脉络,有效的提升了目标内容的展示效果。
另外,目标内容可以通过悬浮窗口的形式显示在当前界面上。
示例性的,在将文档内容录入至资源库后,可以对该文档内容进行章节、页码、段落、句子划分。比如,可以根据文档内容的关键字(例如“第三章XXX”)获取章节信息;可以根据文档内容的换行符获取段落信息;可以根据文档内容的页码获取页码信息;可以根据文档内容的标点符号(“。!?……”等)获取句子信息。而且,获取到的每个句子的章节编号、页码编号、段落编号和句子编号,可以都是从1开始依次增加。
其中,文档内容的章节、页码、段落、句子等编号信息,可以作为句子的标识信息,且每个句子的标识信息具有唯一性。如图2所示,标识信息为“第1章,第15页,第78段,第235句”在文档内容中对应的语句是“你来干什么?”,标识信息为“第1章,第15页,第80段,第237句”在文档内容中对应的语句是“小明已经从学校回到家了。”。
需要注意的是,章节、页码、段落、句子均是从1开始增加,不重新计数;页码、段落、句子编号不因章节编号而重新计数;段落不因章节、页码变化而重新计数;句子不因章节、页码、段落变化而重新计数。
另外,可以针对每个句子进行命名实体识别,包括人名识别、地点识别、事件识别、组织机构识别、专有名字识别等。识别结果包括:实体名、实体在句子中的位置、实体类型等。下面以标识信息为“第1章,第15页,第79段,第236句”对应的句子“小明、小李和小王是同班同学。”为例对命 名实体进行识别,可以得到:实体类型-人物位置-1,实体类型-人物位置-4,实体类型-人物位置7。识别到的命名实体可以存储在数据库模块14中,比如可以以表1的形式存储识别到的命名实体。
表1
而且,可以统计文档内容中人物、地点、位置等实体出现的次数,并可以进行倒序排序,并存储数据库中。其中,统计到的实体信息,可以以表2的形式进行存储。
表2
另外,可以根据句子的标识信息,统计每章、每页的人物、地点、事件等实体出现次数,并存储在数据库中。其中,每章实体出现次数统计可以如表3所示,每页实体出现次数统计可以如表4所示。
表3
表4
从上数统计可知,每章节的实体计数较多,每页实体计数较少,对于某些实体来说,在很多页出现的次数可能为0。
此外,还可以对文档内容进行时序分析。在对文档内容进行时序分析的过程中,可以先将文档内容划分为多个时间窗口,然后对每一时间窗口进行情感分析,将对应的情感分析结果存储,直至所有的时间窗口完成情感分析。
在划分时间窗口的过程中,可以先将连线的X(例如X为3)个段落划分为时间窗口,相邻的两个时间窗口之间位移一个段落,这样就能连续地分析人物之间的情感。根据此方法,可以将文档内容划分为若干个时间窗口。比如,将第n个段落对应第n个窗口。其中,第一个窗口和最后一个窗口对应的时间窗口长度为2(即在X为3的情况下),因为它们分别缺失前段落和后段落。
在情感分析过程中,情感是有方向的,A对B的情感和B对A的情感不 一定一致。而且,分析实体间情感的方法有多种,可以基于深度学习的实体情感分析:将情感分为多级。示例性的,可以将情感分为5级:厌恶(-2)、梳离(-1)、中性(0)、亲近(1)、喜爱(2),并可以由深度学习模型来判断两个实体之间的相互情感属于哪一级。
其中,经情感分析后得到的情感分析结果可以以表5的形式存储在数据库中。
表5
其中,上述过程可以在客户端上执行,也可以在云平台等服务器上执行。
示例性的,用户在通过客户端或者移动终端阅读文档内容时,客户端可以会向服务器发送请求信息,请求信息包括当前页码信息。服务器可以根据接收到的请求信息,向客户端返回识别得到的人物、地点、事件等实体,并在阅读器(可以包括但不限于手机)中通过标识符(色块标识等)标出,然后根据统计到的实体全文计数,并将实体标识符按照计数多少进行着色(比如,按照计数降序排名,将文中实体划分为5等分,计数最高的一份着色最深,计数最少的一份着色最浅)。通过这样设置,可以让用户忽略不重要的配角、地点、事件等实体,提升用户的阅读效率和阅读体验。
需要说明的是,本申请中的目标对象可以是上述实施方式中提及的实体,即目标对象可以是文档内容中的人物、地点、事件等实体信息。
可选地,N大于1;所述显示目标内容,包括:
基于所述行文进度获取M个子文本中的任意一个子文本中的所述N个目标对象之间的关联关系,所述M个子文本是所述文档内容基于预设分隔符划分得到的,M为大于1的整数;
以所述行文进度为时间轴,显示所述N个目标对象在每一个子文本中的关联内容,所述关联内容是基于每一个子文本中的所述N个目标对象之间的关联关系确定的;
其中,所述关联内容包括共现次数、情感级别中的至少一项。
本实施方式中,预设分隔符可以是文档内容的章节分隔符、页码分隔符、段落分隔符或者句子分隔符等,用户可以基于实际需求(比如文档内容的篇幅)合理的设置预设分隔符,以降低划分子文本所需的计算量。
可以理解的是,在本申请中,M个子文本的划分既可以在客户端侧执行,也可以在服务器侧执行。
在N大于1的情况下,即接收到对至少两个目标对象的第一输入的情况下,则可以基于行文进度获取至少两个目标对象的关联关系,以便用户可以基于至少两个目标对象的关联关系,更好地去理解文档内容的发展脉络,提升用户的阅读效率。具体地,可以基于两个目标对象的共现次数、情感级别等关联内容,更好地理解文档内容的发展脉络。
比如,在N为2的情况下,即接收到用户对两个目标对象的第一输入(例如用户同时点击阅读界面上的两个实体词)的情况下,则可以获取两个目标对象在任何时间窗口中的关联关系。
示例性的,在两个目标对象为当前展示内容中的“小明”和“小王”时,则可以得到“小明”和“小王”在任何时间窗口中的情感关系,具体可参见表6。
表6
可选地,所述基于所述行文进度获取M个子文本中的任意一个子文本中的所述N个目标对象之间的关联关系,包括:
在第一子文本为所述文档内容的非首文本的情况下,获取第一子文本中的所述N个目标对象之间的第一关联关系;其中,在所述第一关联关系表征所述N个目标对象之间存在相互关联时,则将所述第一关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;在所述第一关联关系表征所述N个目标对象之间未存在相互关联时,则获取所述N个目标对象在第二子文本中的第二关联关系,并将所述第二关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;所述第一子文本和所述第二子文本为所述M个子文本中的以所述行文进度为时间轴且相邻的子文本,且所述第一子文本的行文顺序位于所述第二子文本之后。
本实施方式中,当第一子文本为文档内容的非首文本的情况下,则可以基于第一子文本中的N个目标对象之间的第一关联关系,确定N个目标对象在第一子文本中的关联关系。
其中,当第一关联关系表征N个目标对象之间存在相互关联时,则可以将第一关联关系确定为N个目标对象在第一子文本中的关联关系;而当第一关联关系表征N个目标对象之间未存在相互关联时,则可以将N个目标对象在第二子文本中的第二关联关系确定为N个目标对象在第一子文本中的关联关系,以便是N个目标对象在文档内容中的关联关系具有连续性,并达到改善用户的阅读体验的目的。
另外,需要说明的是,当第一子文本为文档内容的首文本的情况下,则可以将第一子文本中的N个目标对象之间的关联关系确定为N个目标对象在第一子文本(首文本)的关联关系。
示例性的,在N为2的情况下,两个目标对象可以标记为e1和e2,并标记文档内容在当前界面显示的内容的页码为p,目标对象e1对目标对象e2在第p页的关系值(即关联关系值)标记为R(e1,e2,p);相应地,目标对象e2对目标对象e1在第p页的关系值标记为R(e2,e1,p),文档内容的总页码 数标记为max。文档内容开始前(第1页之前),页码p初始化为0,所有目标对象间的关系值初始化为0,有R(e1,e2,0)=0,R(e2,e1,0)=0。
接下来,进入循环:循环开始,页数增加1,R(e1,e2,p)和R(e2,e1,p),具体地,可以先查询类似表6中的两个目标对象的所有关联关系,并获取e2对e1在第p页的最后一个关系值,作为R(e2,e1,p),若在服务器返回的关联关系中没查到e2对e1在第p页的最后一个关系值,则令R(e2,e1,p)=R(e2,e1,p-1),即认为如果两个目标对象的关联关系没有新的进展,e2对e1的在p页关系值默认与前一页p-1一致。
如果当前页不是最后一页,p增加1,从下一页开始继续循环插值,直到当前页变为最后一页。
由此有R(e2,e1,p)={(0,p=0;R(e2,e1,p-1),当查询不到,0<p<max;查询值,0<p<max),且R(e1,e2,p)有类似结论。如此便完成了实体e1与e2在每一页的相互关系插值。
其中,在插值后,可以在移动终端的显示界面展示目标对象的情感关系的时序分析图。
可选地,所述目标内容为所述目标对象在所述文档内容中的时序图;
其中,所述时序图包括横轴、纵轴和展示曲线,所述横轴以所述行文进度为单位,所述纵轴以所述目标对象在任意一行文单位中的出现次数为单位,所述展示曲线为所述目标对象在每一所述行文单位中的出现次数所连接形成的曲线,且所述行文单位为所述文档内容基于所述行文进度划分得到的。
本实施方式中,时序图可以基于行文进度展示单个目标对象在文档内容中的出现频次,时序图也可以基于行文进度展示至少两个目标对象在文档内容中的关联关系。
可选地,所述时序图中还包括显示在所述横轴上的拉杆图标,且所述拉杆图标在所述横轴上的位置对应当前界面显示的行文单位在所述行文进度上的位置;
所述显示目标内容之后,所述方法还包括:
接收对所述拉杆图标的第二输入;
响应于所述第二输入,确定与所述拉杆图标对应的目标行文单位;
切换至所述目标行文单位中的包括所述目标对象的文本段落并显示。
本实施方式中,通过拉杆图标的设置,用户可以快速的跳转至更新后的拉杆图标对应的目标行文单位中的包括该目标对象的文本段落,以提高用户的阅读效率。
其中,第二输入可以是作用在拉杆图标上的长按操作或者双击操作等。
一些实施方式中,在用户通过手机阅读文档内容时,在接收到用户对手机的显示屏上显示的人物、地点、事件等实体词(目标对象)的点击操作或长按操作的情况下,手机会通过网络向服务器发送请求信息,该请求信息包括了上述点击操作或长按操作所作用的实体名;服务器在接收到该请求信息后,会响应该请求信息,并从服务器的数据库模块中获取该实体在页码、章节中的计数信息,且服务器会将获取到的计数信息发送至手机,以便手机可以基于接收到计数信息生成目标内容。
如图3所示,目标内容可以以悬浮窗口的形式显示在当前界面显示的内容的上方,以展示目标对象(例如人物、地点、事件等实体)在文档内容全文中各个页面和章节的出现频次。如图4所示,目标内容可以是频次时序图,并包括两个坐标轴和一条展示曲线。两个坐标轴中的横轴用于指示页码,且在章节的起始页有标注章节;纵轴用于指示当前页的目标对象的计数。当前页面在全文中的位置可以以拉杆图标(比如直线加圆球的拉杆)的形式展现在该频次时序图中。并且通过拉动拉杆图标,还可以实现调整阅读位置,并直接跳转至拉杆图标所指的页码,即可以跳转至更新后的拉杆图标对应的目标行文单位中的包括该目标对象的文本段落。
另一些实施方式中,目标内容还可以至少两个目标对象的关联关系的时序图。如图5和图6所示,在时序图用于表达至少两个目标对象的关联关系的情况下,该情感时序图以页为横轴(即以文档内容的行文顺序为横轴),并在章节起始页标明章节,以及以至少两个目标对象在每页的情感得分为纵轴。 当前页会以拉杆图标(比如直线加圆球的拉杆)的形式展现在该情感时序图中。而且通过拉动拉杆图标,还可以实现调整阅读位置,并直接跳转至拉杆图标所指的页码,即可以跳转至更新后的拉杆图标对应的目标行文单位中的包括该目标对象的文本段落。
在又一些实施方式中,当第一输入作用于至少两个目标对象的情况下,目标内容不仅可以包括该至少两个目标对象的情感时序图,还可以包括该至少两个目标对象的频次时序图,具体可参见图7。
可选地,所述时序图还包括所述目标对象的图谱图标;
所述显示目标内容之后,所述方法还包括:
接收对所述图谱图标的第三输入;
响应于所述第三输入,显示所述目标对象和关联对象的图谱关系图;
其中,所述关联对象为所述文档内容中与所述目标对象存在关联关系的对象,且所述目标对象和所述关联对象的关联程度与连线长度负相关,所述连线长度为所述目标对象和关联对象在所述图谱关系图中的距离长度。
本实施方式中,在时序图中还可以显示目标对象的图谱图标,通过接收针对图谱图标的第三输入,还可以显示目标对象与其关联对象的图谱关系图,以便用户更好的理解文档内容的发展脉络,实现对文档内容的高效阅读。
其中,第三输入可以是针对图谱图标的长按操作或者双击操作等。
示例性的,在接收到对图4中显示的“查看完整人物图谱”的图谱图标的点击操作时,移动终端会向服务器发送请求信息,该请求信息为图4中显示目标对象(并标记为“中心实体”),以获取该目标对象截止至当前页的图谱底层数据。
该图谱底层数据包括:获取到的实体识别数据,并可以要求获得所有与该中心实体发生段内共现的实体;以及获取到的情感关系数据,并可以要求获得所有与该中心实体发生段内共现的实体与该中心实体的关系数据。
在获取到图谱底层数据后,可以基于该图谱底层数据计算当前页的图谱数据,进而生成图谱关系图。其中,在计算当前页的图谱数据的过程中,主 要分为两步,一是计算截止当前页与中心实体产生段落共现的所有实体及其共现次数,二是可以按照前述内容中所提及的插值法,计算到当前页为止中心实体与其共现实体之间的关系。
示例性的,本实施方式中的图谱关系图可以如图8所示。如图8所示,目标对象(中心实体)“小明”与其他实体的关系在图中可以分为强、中、弱三个级别,且所有其他实体的位置可以参考如图8所示的方式。而且,实体和实体之间的情感级别也可以用连线标识出来,以便用户查看文档内容的整体脉络。
比如,中心实体e0可以置于图谱关系图的中心位置,并可以设置图谱关系图的长边长为L;强关系区域可以设置为以中心实体为圆心,以L/6为半径的圆;中关系区域可以设置为以中心实体为圆心,以L/3为半径的圆;弱关系区域可以设置为强关系区域和中关系区域之外的区域。其中,共现次数最少的实体e1可以随机放置在弱关系区域,与中心实体e0的举例为L1;并可以设置e1和e0的共现次数为C1,其他实体ex和e0的共现次数分别为Cx,ex与中心实体的举例可以设置为Lx=L1*(C1/Cx);此外,还可以限定ex与中心实体的举例Lx,均匀放置ex于图谱关系图中。
进一步地,通过点击图8中的任何实体,都可以将中心实体设定为被点击的实体。比如,通过点击图8中的“小王”,即可以将图8所示的以“小明”为中心实体的图谱关系图切换为如图9所示的以“小王”为中心实体的图谱关系图,即可以获得“小王”的完整人物图谱。
另外,在图谱关系图中,还显示有拉杆图标,通过拉动或点击拉杆图谱,可以调整后的图谱关系图。比如,将图9中的拉杆图标拉动至第六章,此时更新后的图谱关系图如图10所示,在图10所示的图谱关系图中,“小王”和“小李”也熟悉了,情感关系也变为了相互亲近;而“小王”对“小赵”却疏远了,情感关系变为排斥。
通过本申请实施例提供的方案,能够实现文档内容中的实体的重要性计算,并且能够以章节/页码等为时间刻度进行分章节/页码的实体与关系计算。 这样用户在阅读时,可以即时得到实体的重要性信息,并可以选择性的阅读而不影响对文档内容的理解。并且还可以根据对人物等实体的兴趣,拉动拉杆图标,随时调整阅读位置,实现选择性的阅读。另外,通过显示带有时序特征的图谱关系图,还可以让用户轻松的理解文档内容的结构及文档内容中的实体关系,有效的提升了用户的阅读效率。
需要说明的是,移动终端不仅可以是手机、平板等终端设备,还可以是VR设备,即在VR场景下也可以实现上述方案。
进一步地,还可以对文档内容进行领域或类别划分,改善对文档内容的识别分析过程。
另外,还可以结合文档内容中的阅读批注、部分实体的网络百科等内容,以辅助用户阅读,以便帮助用户更好地理解文档内容的发展脉络,实现高效阅读。
本申请实施例的文档显示方法,通过接接收用户对当前界面显示的文档内容中的N个目标对象的第一输入,N为正整数;响应于所述第一输入,显示目标内容,所述目标内容是基于所述文档内容的行文进度和所述N个目标对象生成得到的;其中,所述行文进度用于表征所述文档内容所记载的字符和字符之间的行文顺序。这样用户通过查看显示的目标内容,就可以知晓目标对象在文档内容中的重要性,而且由于目标内容是以行文进度为时间脉络来展示目标对象的,因此通过查看目标内容还可以帮助用户理解文档内容的发展脉络,有效的提升了目标内容的展示效果。
本申请实施例提供的文档显示方法,执行主体可以为文档显示装置。本申请实施例中以文档显示装置执行文档显示方法为例,说明本申请实施例提供的文档显示装置。
参见图11,图11是本申请一实施例提供的文档显示装置的结构图,如图11所示,该装置1100包括:
第一接收模块1101,用于接收用户对当前界面显示的文档内容中的N个目标对象的第一输入,N为正整数;
第一显示模块1102,用于响应于所述第一输入,显示目标内容,所述目标内容是基于所述文档内容的行文进度和所述N个目标对象生成得到的;
其中,所述行文进度用于表征所述文档内容所记载的字符和字符之间的行文顺序。
可选地,N大于1;所述第一显示模块1102包括:
获取单元,用于基于所述行文进度获取M个子文本中的任意一个子文本中的所述N个目标对象之间的关联关系,所述M个子文本是所述文档内容基于预设分隔符划分得到的,M为大于1的整数;
显示单元,用于以所述行文进度为时间轴,显示所述N个目标对象在每一个子文本中的关联内容,所述关联内容是基于每一个子文本中的所述N个目标对象之间的关联关系确定的;
其中,所述关联内容包括共现次数、情感级别中的至少一项。
可选地,所述获取单元,具体用于在第一子文本为所述文档内容的非首文本的情况下,获取第一子文本中的所述N个目标对象之间的第一关联关系;其中,在所述第一关联关系表征所述N个目标对象之间存在相互关联时,则将所述第一关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;在所述第一关联关系表征所述N个目标对象之间未存在相互关联时,则获取所述N个目标对象在第二子文本中的第二关联关系,并将所述第二关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;所述第一子文本和所述第二子文本为所述M个子文本中的以所述行文进度为时间轴且相邻的子文本,且所述第一子文本的行文顺序位于所述第二子文本之后。
可选地,所述目标内容为所述目标对象在所述文档内容中的时序图;
其中,所述时序图包括横轴、纵轴和展示曲线,所述横轴以所述行文进度为单位,所述纵轴以所述目标对象在任意一行文单位中的出现次数为单位,所述展示曲线为所述目标对象在每一所述行文单位中的出现次数所连接形成的曲线,且所述行文单位为所述文档内容基于所述行文进度划分得到的。
可选地,所述时序图中还包括显示在所述横轴上的拉杆图标,且所述拉 杆图标在所述横轴上的位置对应当前界面显示的行文单位在所述行文进度上的位置;
所述装置1100还包括:
第二接收模块,用于接收对所述拉杆图标的第二输入;
确定模块,用于响应于所述第二输入,确定与所述拉杆图标对应的目标行文单位;
切换显示模块,用于切换至所述目标行文单位中的包括所述目标对象的文本段落并显示。
可选地,所述时序图还包括所述目标对象的图谱图标;
所述装置1100还包括:
第三接收模块,用于接收对所述图谱图标的第三输入;
第二显示模块,用于响应于所述第三输入,显示所述目标对象和关联对象的图谱关系;
其中,所述关联对象为所述文档内容中与所述目标对象存在关联关系的对象,且所述目标对象和所述关联对象的关联程度与连线长度负相关,所述连线长度为所述目标对象和关联对象在所述图谱关系中的距离长度。
本申请实施例中的文档显示装置可以是电子设备,也可以是电子设备中的部件,例如集成电路或芯片。该电子设备可以是终端,也可以为除终端之外的其他设备。示例性的,电子设备可以为手机、平板电脑、笔记本电脑、掌上电脑、车载电子设备、移动上网装置(Mobile Internet Device,MID)、增强现实(Augmented Reality,AR)/虚拟现实(Virtual Reality,VR)设备、机器人、可穿戴设备、超级移动个人计算机(Ultra-Mobile Personal Computer,UMPC)、上网本或者个人数字助理(Personal Digital Assistant,PDA)等,还可以为网络附属存储器(Network Attached Storage,NAS)、个人计算机(Personal Computer,PC)、电视机(Television,TV)、柜员机或者自助机等,本申请实施例不作具体限定。
本申请实施例中的文档显示装置可以为具有操作系统的装置。该操作系 统可以为安卓(Android)操作系统,可以为ios操作系统,还可以为其他可能的操作系统,本申请实施例不作具体限定。
本申请实施例提供的文档显示装置能够实现图1的方法实施例实现的各个过程,为避免重复,这里不再赘述。
可选地,如图12所示,本申请实施例还提供一种电子设备1200,包括处理器1201和存储器1202,存储器1202上存储有可在所述处理器1201上运行的程序或指令,该程序或指令被处理器1201执行时实现上述文档显示方法实施例的各个步骤,且能达到相同的技术效果,为避免重复,这里不再赘述。
需要说明的是,本申请实施例中的电子设备包括上述所述的移动电子设备和非移动电子设备。
参见图13,图13是本申请一实施例提供的电子设备的结构图,如图13所示,该电子设备1300包括但不限于:射频单元1301、网络模块1302、音频输出单元1303、输入单元1304、传感器1305、显示单元1306、用户输入单元1307、接口单元1308、存储器1309、以及处理器1310等部件。
本领域技术人员可以理解,电子设备1300还可以包括给各个部件供电的电源(比如电池),电源可以通过电源管理系统与处理器1310逻辑相连,从而通过电源管理系统实现管理充电、放电、以及功耗管理等功能。图13中示出的电子设备结构并不构成对电子设备的限定,电子设备可以包括比图示更多或更少的部件,或者组合某些部件,或者不同的部件布置,在此不再赘述。
其中,用户输入单元1307,用于接收用户对当前界面显示的文档内容中的N个目标对象的第一输入,N为正整数,N为正整数;
显示单元1306,用于响应于所述第一输入,显示目标内容,所述目标内容是基于所述文档内容的行文进度和所述N个目标对象生成得到的;
其中,所述行文进度用于表征所述文档内容所记载的字符和字符之间的行文顺序。
可选地,N大于1;处理器1310,用于基于所述行文进度获取M个子文 本中的任意一个子文本中的所述N个目标对象之间的关联关系,所述M个子文本是所述文档内容基于预设分隔符划分得到的,M为大于1的整数;显示单元1306,用于以所述行文进度为时间轴,显示所述N个目标对象在每一个子文本中的关联内容,所述关联内容是基于每一个子文本中的所述N个目标对象之间的关联关系确定的;其中,所述关联内容包括共现次数、情感级别中的至少一项。
可选地,处理器1310,用于在第一子文本为所述文档内容的非首文本的情况下,获取第一子文本中的所述N个目标对象之间的第一关联关系;其中,在所述第一关联关系表征所述N个目标对象之间存在相互关联时,则将所述第一关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;在所述第一关联关系表征所述N个目标对象之间未存在相互关联时,则获取所述N个目标对象在第二子文本中的第二关联关系,并将所述第二关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;所述第一子文本和所述第二子文本为所述M个子文本中的以所述行文进度为时间轴且相邻的子文本,且所述第一子文本的行文顺序位于所述第二子文本之后。
可选地,所述目标内容为所述目标对象在所述文档内容中的时序图;
其中,所述时序图包括横轴、纵轴和展示曲线,所述横轴以所述行文进度为单位,所述纵轴以所述目标对象在任意一行文单位中的出现次数为单位,所述展示曲线为所述目标对象在每一所述行文单位中的出现次数所连接形成的曲线,且所述行文单位为所述文档内容基于所述行文进度划分得到的。
可选地,所述时序图中还包括显示在所述横轴上的拉杆图标,且所述拉杆图标在所述横轴上的位置对应当前界面显示的行文单位在所述行文进度上的位置;用户输入单元1307,用于接收对所述拉杆图标的第二输入;处理器1310,用于响应于所述第二输入,确定与所述拉杆图标对应的目标行文单位;显示单元1306,用于切换至所述目标行文单位中的包括所述目标对象的文本段落并显示。
可选地,所述时序图还包括所述目标对象的图谱图标;用户输入单元1307, 用于接收对所述图谱图标的第三输入;显示单元1306,用于响应于所述第三输入,显示所述目标对象和关联对象的图谱关系;其中,所述关联对象为所述文档内容中与所述目标对象存在关联关系的对象,且所述目标对象和所述关联对象的关联程度与连线长度负相关,所述连线长度为所述目标对象和关联对象在所述图谱关系中的距离长度。
应理解的是,本申请实施例中,输入单元1304可以包括图形处理器(Graphics Processing Unit,GPU)13041和麦克风13042,图形处理器13041对在视频捕获模式或图像捕获模式中由图像捕获装置(如摄像头)获得的静态图片或视频的图像数据进行处理。显示单元1306可包括显示面板13061,可以采用液晶显示器、有机发光二极管等形式来配置显示面板13061。用户输入单元1307包括触控面板13071以及其他输入设备13072中的至少一种。触控面板13071,也称为触摸屏。触控面板13071可包括触摸检测装置和触摸控制器两个部分。其他输入设备13072可以包括但不限于物理键盘、功能键(比如音量控制按键、开关按键等)、轨迹球、鼠标、操作杆,在此不再赘述。
存储器1309可用于存储软件程序以及各种数据。存储器1309可主要包括存储程序或指令的第一存储区和存储数据的第二存储区,其中,第一存储区可存储操作系统、至少一个功能所需的应用程序或指令(比如声音播放功能、图像播放功能等)等。此外,存储器1309可以包括易失性存储器或非易失性存储器,或者,存储器1309可以包括易失性和非易失性存储器两者。其中,非易失性存储器可以是只读存储器(Read-Only Memory,ROM)、可编程只读存储器(Programmable ROM,PROM)、可擦除可编程只读存储器(Erasable PROM,EPROM)、电可擦除可编程只读存储器(Electrically EPROM,EEPROM)或闪存。易失性存储器可以是随机存取存储器(Random Access Memory,RAM),静态随机存取存储器(Static RAM,SRAM)、动态随机存取存储器(Dynamic RAM,DRAM)、同步动态随机存取存储器(Synchronous DRAM,SDRAM)、双倍数据速率同步动态随机存取存储器(Double Data Rate SDRAM, DDRSDRAM)、增强型同步动态随机存取存储器(Enhanced SDRAM,ESDRAM)、同步连接动态随机存取存储器(Synch link DRAM,SLDRAM)和直接内存总线随机存取存储器(Direct Rambus RAM,DRRAM)。本申请实施例中的存储器1309包括但不限于这些和任意其它适合类型的存储器。
处理器1310可包括一个或多个处理单元;可选地,处理器1310集成应用处理器和调制解调处理器,其中,应用处理器主要处理涉及操作系统、用户界面和应用程序等的操作,调制解调处理器主要处理无线通信信号,如基带处理器。可以理解的是,上述调制解调处理器也可以不集成到处理器1310中。
本申请实施例还提供一种可读存储介质,该存储介质可以是易失的或非易失的,所述可读存储介质上存储有程序或指令,该程序或指令被处理器执行时实现上述文档显示方法实施例的各个过程,且能达到相同的技术效果,为避免重复,这里不再赘述。
其中,所述处理器为上述实施例中所述的电子设备中的处理器。所述可读存储介质,包括计算机可读存储介质,如计算机只读存储器ROM、随机存取存储器RAM、磁碟或者光盘等。
本申请实施例另提供了一种芯片,所述芯片包括处理器和通信接口,所述通信接口和所述处理器耦合,所述处理器用于运行程序或指令,实现上述文档显示方法实施例的各个过程,且能达到相同的技术效果,为避免重复,这里不再赘述。
应理解,本申请实施例提到的芯片还可以称为系统级芯片、系统芯片、芯片系统或片上系统芯片等。
本申请实施例提供一种计算机程序产品,该程序产品被存储在存储介质中,该程序产品被至少一个处理器执行以实现如上述文档显示方法实施例的各个过程,且能达到相同的技术效果,为避免重复,这里不再赘述。
需要说明的是,在本文中,术语“包括”、“包含”或者其任何其他变体意在涵盖非排他性的包含,从而使得包括一系列要素的过程、方法、物品或 者装置不仅包括那些要素,而且还包括没有明确列出的其他要素,或者是还包括为这种过程、方法、物品或者装置所固有的要素。在没有更多限制的情况下,由语句“包括一个……”限定的要素,并不排除在包括该要素的过程、方法、物品或者装置中还存在另外的相同要素。此外,需要指出的是,本申请实施方式中的方法和装置的范围不限按示出或讨论的顺序来执行功能,还可包括根据所涉及的功能按基本同时的方式或按相反的顺序来执行功能,例如,可以按不同于所描述的次序来执行所描述的方法,并且还可以添加、省去、或组合各种步骤。另外,参照某些示例所描述的特征可在其他示例中被组合。
通过以上的实施方式的描述,本领域的技术人员可以清楚地了解到上述实施例方法可借助软件加必需的通用硬件平台的方式来实现,当然也可以通过硬件,但很多情况下前者是更佳的实施方式。基于这样的理解,本申请的技术方案本质上或者说对相关技术做出贡献的部分可以以计算机产品的形式体现出来,该计算机软件产品存储在一个存储介质(如ROM/RAM、磁碟、光盘)中,包括若干指令用以使得一台终端(可以是手机,计算机,服务器,或者网络设备等)执行本申请各个实施例所述的方法。
上面结合附图对本申请的实施例进行了描述,但是本申请并不局限于上述的具体实施方式,上述的具体实施方式仅仅是示意性的,而不是限制性的,本领域的普通技术人员在本申请的启示下,在不脱离本申请宗旨和权利要求所保护的范围情况下,还可做出很多形式,均属于本申请的保护之内。

Claims (17)

  1. 一种文档显示方法,包括:
    接收用户对当前界面显示的文档内容中的N个目标对象的第一输入,N为正整数;
    响应于所述第一输入,显示目标内容,所述目标内容是基于所述文档内容的行文进度和所述N个目标对象生成得到的;
    其中,所述行文进度用于表征所述文档内容所记载的字符和字符之间的行文顺序。
  2. 根据权利要求1所述的方法,其中,N大于1;所述显示目标内容,包括:
    基于所述行文进度获取M个子文本中的任意一个子文本中的所述N个目标对象之间的关联关系,所述M个子文本是所述文档内容基于预设分隔符划分得到的,M为大于1的整数;
    以所述行文进度为时间轴,显示所述N个目标对象在每一个子文本中的关联内容,所述关联内容是基于每一个子文本中的所述N个目标对象之间的关联关系确定的;
    其中,所述关联内容包括共现次数、情感级别中的至少一项。
  3. 根据权利要求2所述的方法,其中,所述基于所述行文进度获取M个子文本中的任意一个子文本中的所述N个目标对象之间的关联关系,包括:
    在第一子文本为所述文档内容的非首文本的情况下,获取第一子文本中的所述N个目标对象之间的第一关联关系;其中,在所述第一关联关系表征所述N个目标对象之间存在相互关联时,则将所述第一关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;在所述第一关联关系表征所述N个目标对象之间未存在相互关联时,则获取所述N个目标对象在第二子文本中的第二关联关系,并将所述第二关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;所述第一子文本和所述第二子文本为所述M 个子文本中的以所述行文进度为时间轴且相邻的子文本,且所述第一子文本的行文顺序位于所述第二子文本之后。
  4. 根据权利要求1所述的方法,其中,所述目标内容为所述目标对象在所述文档内容中的时序图;
    其中,所述时序图包括横轴、纵轴和展示曲线,所述横轴以所述行文进度为单位,所述纵轴以所述目标对象在任意一行文单位中的出现次数为单位,所述展示曲线为所述目标对象在每一所述行文单位中的出现次数所连接形成的曲线,且所述行文单位为所述文档内容基于所述行文进度划分得到的。
  5. 根据权利要求4所述的方法,其中,所述时序图中还包括显示在所述横轴上的拉杆图标,且所述拉杆图标在所述横轴上的位置对应当前界面显示的行文单位在所述行文进度上的位置;
    所述显示目标内容之后,所述方法还包括:
    接收对所述拉杆图标的第二输入;
    响应于所述第二输入,确定与所述拉杆图标对应的目标行文单位;
    切换至所述目标行文单位中的包括所述目标对象的文本段落并显示。
  6. 根据权利要求4所述的方法,其中,所述时序图还包括所述目标对象的图谱图标;
    所述显示目标内容之后,所述方法还包括:
    接收对所述图谱图标的第三输入;
    响应于所述第三输入,显示所述目标对象和关联对象的图谱关系;
    其中,所述关联对象为所述文档内容中与所述目标对象存在关联关系的对象,且所述目标对象和所述关联对象的关联程度与连线长度负相关,所述连线长度为所述目标对象和关联对象在所述图谱关系中的距离长度。
  7. 一种文档显示装置,包括:
    第一接收模块,用于接收用户对当前界面显示的文档内容中的N个目标对象的第一输入,N为正整数;
    第一显示模块,用于响应于所述第一输入,显示目标内容,所述目标内 容是基于所述文档内容的行文进度和所述N个目标对象生成得到的;
    其中,所述行文进度用于表征所述文档内容所记载的字符和字符之间的行文顺序。
  8. 根据权利要求7所述的装置,其中,N大于1;所述第一显示模块包括:
    获取单元,用于基于所述行文进度获取M个子文本中的任意一个子文本中的所述N个目标对象之间的关联关系,所述M个子文本是所述文档内容基于预设分隔符划分得到的,M为大于1的整数;
    显示单元,用于以所述行文进度为时间轴,显示所述N个目标对象在每一个子文本中的关联内容,所述关联内容是基于每一个子文本中的所述N个目标对象之间的关联关系确定的;
    其中,所述关联内容包括共现次数、情感级别中的至少一项。
  9. 根据权利要求8所述的装置,其中,所述获取单元,具体用于在第一子文本为所述文档内容的非首文本的情况下,获取第一子文本中的所述N个目标对象之间的第一关联关系;其中,在所述第一关联关系表征所述N个目标对象之间存在相互关联时,则将所述第一关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;在所述第一关联关系表征所述N个目标对象之间未存在相互关联时,则获取所述N个目标对象在第二子文本中的第二关联关系,并将所述第二关联关系确定为所述N个目标对象在所述第一子文本中的关联关系;所述第一子文本和所述第二子文本为所述M个子文本中的以所述行文进度为时间轴且相邻的子文本,且所述第一子文本的行文顺序位于所述第二子文本之后。
  10. 根据权利要求7所述的装置,其中,所述目标内容为所述目标对象在所述文档内容中的时序图;
    其中,所述时序图包括横轴、纵轴和展示曲线,所述横轴以所述行文进度为单位,所述纵轴以所述目标对象在任意一行文单位中的出现次数为单位,所述展示曲线为所述目标对象在每一所述行文单位中的出现次数所连接形成 的曲线,且所述行文单位为所述文档内容基于所述行文进度划分得到的。
  11. 根据权利要求10所述的装置,其中,所述时序图中还包括显示在所述横轴上的拉杆图标,且所述拉杆图标在所述横轴上的位置对应当前界面显示的行文单位在所述行文进度上的位置;
    所述装置还包括:
    第二接收模块,用于接收对所述拉杆图标的第二输入;
    确定模块,用于响应于所述第二输入,确定与所述拉杆图标对应的目标行文单位;
    切换显示模块,用于切换至所述目标行文单位中的包括所述目标对象的文本段落并显示。
  12. 根据权利要求10所述的装置,其中,所述时序图还包括所述目标对象的图谱图标;
    所述装置还包括:
    第三接收模块,用于接收对所述图谱图标的第三输入;
    第二显示模块,用于响应于所述第三输入,显示所述目标对象和关联对象的图谱关系;
    其中,所述关联对象为所述文档内容中与所述目标对象存在关联关系的对象,且所述目标对象和所述关联对象的关联程度与连线长度负相关,所述连线长度为所述目标对象和关联对象在所述图谱关系中的距离长度。
  13. 一种电子设备,包括处理器和存储器,所述存储器上存储有可在所述处理器上运行的程序或指令,所述程序或指令被所述处理器执行时实现如权利要求1至6中任一项所述的文档显示方法的步骤。
  14. 一种可读存储介质,所述可读存储介质上存储程序或指令,所述程序或指令被处理器执行时实现如权利要求1至6中任一项所述的文档显示方法的步骤。
  15. 一种芯片,所述芯片包括处理器和通信接口,所述通信接口和所述处理器耦合,所述处理器用于运行程序或指令,实现如权利要求1至6中任 一项所述的文档显示方法的步骤。
  16. 一种计算机程序产品,所述计算机程序产品被存储在非易失的存储介质中,所述计算机程序产品被至少一个处理器执行以实现如权利要求1至6中任一项所述的文档显示方法的步骤。
  17. 一种电子设备,被配置为执行如权利要求1至6中任一项所述的文档显示方法的步骤。
PCT/CN2024/071031 2023-01-13 2024-01-08 文档显示方法、装置及电子设备 Ceased WO2024149183A1 (zh)

Applications Claiming Priority (2)

Application Number Priority Date Filing Date Title
CN202310068177.4 2023-01-13
CN202310068177.4A CN116149528A (zh) 2023-01-13 2023-01-13 文档显示方法、装置及电子设备

Publications (1)

Publication Number Publication Date
WO2024149183A1 true WO2024149183A1 (zh) 2024-07-18

Family

ID=86352195

Family Applications (1)

Application Number Title Priority Date Filing Date
PCT/CN2024/071031 Ceased WO2024149183A1 (zh) 2023-01-13 2024-01-08 文档显示方法、装置及电子设备

Country Status (2)

Country Link
CN (1) CN116149528A (zh)
WO (1) WO2024149183A1 (zh)

Families Citing this family (1)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
CN116149528A (zh) * 2023-01-13 2023-05-23 维沃移动通信有限公司 文档显示方法、装置及电子设备

Citations (5)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US20180004397A1 (en) * 2016-06-29 2018-01-04 Google Inc. Systems and Methods of Providing Content Selection
CN112051946A (zh) * 2020-08-06 2020-12-08 北京达佳互联信息技术有限公司 文档数据的展示方法、装置、系统、电子设备及存储介质
CN112417059A (zh) * 2020-10-30 2021-02-26 联想(北京)有限公司 一种信息展示方法、装置及设备
CN113326357A (zh) * 2021-08-03 2021-08-31 北京明略软件系统有限公司 文档信息的查阅方法、装置、电子设备和计算机可读介质
CN116149528A (zh) * 2023-01-13 2023-05-23 维沃移动通信有限公司 文档显示方法、装置及电子设备

Patent Citations (5)

* Cited by examiner, † Cited by third party
Publication number Priority date Publication date Assignee Title
US20180004397A1 (en) * 2016-06-29 2018-01-04 Google Inc. Systems and Methods of Providing Content Selection
CN112051946A (zh) * 2020-08-06 2020-12-08 北京达佳互联信息技术有限公司 文档数据的展示方法、装置、系统、电子设备及存储介质
CN112417059A (zh) * 2020-10-30 2021-02-26 联想(北京)有限公司 一种信息展示方法、装置及设备
CN113326357A (zh) * 2021-08-03 2021-08-31 北京明略软件系统有限公司 文档信息的查阅方法、装置、电子设备和计算机可读介质
CN116149528A (zh) * 2023-01-13 2023-05-23 维沃移动通信有限公司 文档显示方法、装置及电子设备

Also Published As

Publication number Publication date
CN116149528A (zh) 2023-05-23

Similar Documents

Publication Publication Date Title
CN103135884A (zh) 以圈选方式进行检索的输入方法、系统及其装置
WO2022052817A1 (zh) 搜索处理方法、装置、终端及存储介质
CN112631437A (zh) 信息推荐方法、装置及电子设备
WO2024036616A1 (zh) 一种基于终端的问答方法及装置
CN113407828B (zh) 一种搜索方法、装置和用于搜索的装置
CN113157753B (zh) 显示方法、装置及电子设备
CN112882623B (zh) 文本处理方法、装置、电子设备及存储介质
WO2021254251A1 (zh) 输入显示方法、装置及电子设备
CN115437736A (zh) 一种笔记记录方法和装置
WO2020056948A1 (zh) 一种数据处理方法、装置和用于数据处理的装置
WO2024149183A1 (zh) 文档显示方法、装置及电子设备
CN116910368A (zh) 内容处理方法、装置、设备和存储介质
CN108197105A (zh) 自然语言处理方法、装置、存储介质及电子设备
CN113157966B (zh) 显示方法、装置及电子设备
WO2022166811A1 (zh) 信息处理方法、装置、电子设备和存储介质
WO2024179519A1 (zh) 语义识别方法及其装置
CN111913563B (zh) 一种基于半监督学习的人机交互方法及装置
WO2025055822A1 (zh) 粘贴方法、装置和电子设备
CN116152827B (zh) 题目作答方法、装置、设备及介质
WO2024255733A1 (zh) 搜索方法、装置、设备及介质
WO2024160133A1 (zh) 图像生成方法、装置、电子设备和存储介质
US20260105090A1 (en) Conversational computing device assistant
CN113641635B (zh) 文件排序方法、文件排序装置、电子设备和存储介质
CN112732464B (zh) 粘贴方法、装置及电子设备
CN119576461A (zh) 界面显示方法、装置、电子设备及存储介质

Legal Events

Date Code Title Description
121 Ep: the epo has been informed by wipo that ep was designated in this application

Ref document number: 24741182

Country of ref document: EP

Kind code of ref document: A1

NENP Non-entry into the national phase

Ref country code: DE

32PN Ep: public notification in the ep bulletin as address of the adressee cannot be established

Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205A DATED 17.11.2025)