WO2018014543A1 - 信息查询方法、信息查询装置、存储介质及终端 - Google Patents
信息查询方法、信息查询装置、存储介质及终端 Download PDFInfo
- Publication number
- WO2018014543A1 WO2018014543A1 PCT/CN2017/073389 CN2017073389W WO2018014543A1 WO 2018014543 A1 WO2018014543 A1 WO 2018014543A1 CN 2017073389 W CN2017073389 W CN 2017073389W WO 2018014543 A1 WO2018014543 A1 WO 2018014543A1
- Authority
- WO
- WIPO (PCT)
- Prior art keywords
- information
- rumor
- queried
- similarity
- rumbling
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Ceased
Links
Classifications
-
- G—PHYSICS
- G06—COMPUTING OR CALCULATING; COUNTING
- G06F—ELECTRIC DIGITAL DATA PROCESSING
- G06F16/00—Information retrieval; Database structures therefor; File system structures therefor
Definitions
- Invention name information inquiry method, information inquiry device, storage medium, and terminal
- the present invention relates to the field of information technology, and in particular, to an information query method, an information query device, a storage medium, and a terminal.
- a rumor engine emerged in the form of a rumored website.
- the user can enter the keyword of the information in the rumor website. To check the authenticity of this information.
- the present invention provides an information query method, an information query device, a storage medium, and a terminal, which are used to improve the reliability of a spoof query.
- An aspect of the present invention provides an information query method, including:
- the performing a matching query based on the to-be-queried information in a preset rumor database includes: Performing similarity calculations on the to-be-queried information and each piece of plagiar information in the rumor database by using M matching algorithms, and obtaining M similarities between the to-be-queried information and the pieces of rumbling information. a score, where the M is greater than or equal to 2;
- the plagiar information of the similarity total score higher than the preset threshold is determined as the matching rumbling information.
- the obtaining a total score of the similarity between the information to be queried and the pieces of placard information further includes: [0015] when there is no plagiarism information whose similarity total score is higher than the threshold, the plagiar information is sorted according to the order of the similarity total scores from high to low;
- the information is specifically a document title
- the rumor database further includes: a document link address associated with the rumbling information
- the matching query is performed based on the to-be-queried information in the preset rumor database, and further comprising: displaying a document link address associated with the plagiar information to be displayed.
- the triggering of the rumor query instruction acquires information selected by the current cursor as information to be queried.
- a second aspect of the present invention provides an information query apparatus, including:
- a query unit configured to perform a matching query based on the to-be-queried information in a preset rumor database, where the rumor database includes: plagiar information captured from two or more rum;
- a display unit configured to: when the query unit queries the matching plagiarism information, display the matching ⁇ Information.
- the query unit includes: [0027] a similarity score calculation unit, configured to separately perform the check by using M matching algorithms Performing a similarity calculation with each piece of plagiar information in the rumor database, and obtaining M similarity scores of the information to be queried and the pieces of rumbling information, wherein the natural number of the M is greater than or equal to 2 ;
- a weighted summation unit configured to perform weighted summation on the M similarity scores of the to-be-queried information and the pieces of plagiar information, respectively, according to weights set for the various matching algorithms, Obtaining a total score of similarity between the information to be queried and the pieces of rumbling information;
- a determining unit configured to determine the plagiar information that the similarity total score is higher than a preset threshold as a matching ⁇ ⁇ .
- the information querying apparatus further includes:
- a sorting unit configured to: when there is no plagiarism information whose similarity total score is higher than the threshold, sort the placards according to a ranking of the similarity total scores from high to low;
- the rumor information is specifically a document title, and the rumor database further includes: a document link address associated with the rumbling information;
- the display unit is further configured to: display a document link address associated with the rumbling information to be displayed.
- the obtaining unit includes:
- an instruction receiving unit configured to receive a rumor query instruction
- a sub-acquisition unit configured to acquire information selected by the current cursor as information to be queried, triggered by the rumor query instruction received by the instruction receiving unit.
- a third aspect of the present invention provides a storage medium storing one or more passes
- the one or more programs may be executed by one or more processors for:
- the plagiar information of the similarity total score higher than the preset threshold is determined as the matching rumbling information.
- the plagiar information is sorted according to the order of the similarity total scores from high to low;
- the information is specifically a document title
- the rumor database further includes: a document link address associated with the rumbling information
- the one or more programs may also be executed by the one or more processors for: displaying The link address of the document associated with the rumbling information to be displayed.
- a fourth aspect of the present invention provides a terminal, where the terminal includes: at least one processor, at least one input device, at least one output device, and a memory;
- the memory is configured to store an instruction
- the processor is configured to execute the instruction stored by the memory
- the processor is configured to:
- the matching rumbling information is queried, the matching rumbling information is displayed by the output device.
- the performing a matching query based on the to-be-queried information in a preset rumor database includes:
- the plagiar information of the similarity total score higher than the preset threshold is determined as the matching rumbling information.
- the processor is further configured to:
- the plagiar information is sorted according to the order of the similarity total scores from high to low;
- the rumbling information is a document title
- the rumor database further includes: a document link address associated with the rumbling information
- the processor is further configured to: after performing a matching query based on the to-be-queried information in a preset rumor database, displaying a document link address associated with the plagiar information to be displayed.
- the acquiring The information to be inquired includes:
- the information selected by the current cursor is obtained as the information to be queried.
- FIG. 1-b is a schematic diagram of an interface for inputting information to be queried according to the present invention.
- FIG. 3 is a schematic structural diagram of an embodiment of a terminal according to the present invention. Embodiments of the invention
- Step 101 Acquire information to be queried
- the information to be queried may be composed of one or two or more keywords, or may be composed of a character string, which is not limited herein.
- the user is provided with a text input control, so that the user inputs the information to be queried through the text input control, and when receiving the instruction that triggers the information query method shown in FIG. 1-a,
- the information entered in the text input control is used as the information to be queried.
- the user can input the information to be queried through the text input control 11 (for example, "the shooting incident occurred in the XX area"), and then click the button named "query”. 12 triggers the information query method shown in FIG. 1-a, and step 101 acquires the "shot shooting event in the XX area" input by the current text input control 11 as the information to be queried.
- menu 21 (For example, long press on the selected area) Activate menu 21. From the menu 21, click the option named "Query" to enter the above rumor query command.
- the information to be queried may be obtained in other manners in the embodiment of the present invention, which is not limited herein.
- Step 102 Perform a matching query based on the to-be-queried information in a preset rumor database; [0087] wherein, the rumor database includes: plagiarism information captured from two or more rumor.
- a rumor database that composes information from a plurality of rumors websites constitutes a big data, and a rumor engine is formed based on the rumor database.
- the plagiar information may be a document title
- the information captured from the multiple rumors may be based on the document title and the corresponding document link address.
- the rumbling information may also be a specific document content, where Not limited.
- step 102 may perform a matching query by using a Chinese language parsing method, and step 102 includes: using M matching algorithms to respectively perform the similarity between the information to be queried and each plagiar information in the rumor database. Calculating, obtaining M similarity scores of the information to be queried and each of the rumbling information; and, according to the weights set for the foregoing various matching algorithms, respectively, the information to be queried and the rum of the foregoing rumors The similarity scores are weighted and summed, and the total scores of the similarity between the information to be queried and the rumbling information are obtained; the plagiar information with the total score of the similarity higher than the preset threshold is determined as the matching rumbling information.
- step 102 the plagiar information of the total score similarity higher than A' is determined as the matching rumbling information, that is, the second rumor information is Determine the rumor information for matching.
- the M matching algorithms may include, but are not limited to, the following algorithms: a synonym encoding algorithm, a longest substring coincidence algorithm, and a minimum edit distance algorithm.
- a synonym encoding algorithm a longest substring coincidence algorithm
- a minimum edit distance algorithm a minimum edit distance algorithm
- the synonym encoding algorithm is an algorithm based on a hierarchically constructed synonym dictionary.
- the synonym dictionary the closer the closer the words are, the smaller the encoding distance is, and the greater the similarity score.
- Semantic-based synonym analysis can be obtained by obtaining the corresponding coding and calculating the distance.
- the matching query may be performed based on the to-be-queried information in the preset rumor database, for example, using a preset matching algorithm to separately query the information to be queried.
- the similarity calculation is performed on each plagiar information in the rumor database, and the similarity scores of the information to be queried and the rumbling information are obtained, and the plagiar information whose similarity score is higher than another preset threshold is determined as a match.
- the search engine may be used to perform a matching query based on the information to be queried in the rumor database, and the rumor database index is used as a search area, and the information to be queried is used for active search and acquisition.
- the matched rumbling information is displayed for the user to consult.
- the rumbling information is specifically a document title
- the rumor database further includes: a document link address associated with the rumbling information. Then, after step 102, the document link address associated with the rumbling information to be displayed may be further displayed, so that the user can refer to the specific document content by accessing the displayed document link address.
- the total score according to the similarity is The high-to-low order sorts the above-mentioned pieces of placard information, and displays the first N pieces of rumbling information according to the sorted result, wherein the above N is a preset positive integer.
- the information query method in the embodiment of the present invention may be specifically implemented by an information query device, and the information query device may be integrated into the smart terminal in a software form (for example, an APP).
- the above intelligent terminal may specifically be a smart phone, a tablet computer, a personal computer or other electronic terminal, and is not limited herein.
- the present invention when the information to be queried is obtained, a matching query is performed based on the to-be-queried information in a preset rumor database, and the matching information is displayed in the matching query.
- the rumor information because the rumor database is a rumor database of big data composed of rumbling information captured from multiple rum, the invention can provide more comprehensive rumbling information than the traditional rumor engine, thereby effectively Improve the reliability of rumor queries.
- the above-mentioned storage medium may be a read only memory, a magnetic disk or an optical disk or the like.
- An embodiment of the present invention provides an information query apparatus.
- the information querying apparatus 200 in the embodiment of the present invention includes:
- the obtaining unit 201 is configured to obtain information to be queried; [0105]
- the query unit 202 is configured to perform a matching query based on the to-be-queried information in a preset rumor database, where the rumor database includes: plagiar information captured from two or more rum;
- the query unit 202 includes:
- a similarity score calculation unit configured to perform similarity calculation on the to-be-queried information and each piece of plagiar information in the rumor database by using M matching algorithms, and obtain the to-be-queried information and the M similarity scores of each piece of rumbling information, wherein the M is greater than or equal to 2 natural numbers;
- a weighted summation unit configured to perform weighted summation on the M similarity scores of the to-be-queried information and the pieces of plagiar information, respectively, according to weights set for the various matching algorithms, Obtaining a total score of similarity between the information to be queried and the pieces of rumbling information;
- a determining unit configured to determine plagiar information that the similarity total score is higher than a preset threshold as a matching ⁇ ⁇ .
- the rumbling information is specifically a document title
- the rumor database further includes: a document link address associated with the rumbling information
- the display unit 203 is further configured to: display, related to the plagiar information to be displayed Document link address.
- the obtaining unit 201 includes:
- an instruction receiving unit configured to receive a rumor query instruction
- a sub-acquisition unit configured to acquire information selected by the current cursor as information to be queried, triggered by the rumor query instruction received by the instruction receiving unit.
- the information query apparatus in the embodiment of the present invention may be integrated in the smart terminal in a software form (for example, an APP).
- the smart terminal may be a smart phone, a tablet computer, a personal computer or other electronic terminal, which is not limited herein.
- the present invention when the information to be queried is obtained, a matching query is performed based on the to-be-queried information in a preset rumor database, and the matching information is displayed in the matching query.
- the rumor information because the rumor database is a rumor database of big data composed of rumbling information captured from multiple rum, the invention can provide more comprehensive rumbling information than the traditional rumor engine, thereby effectively Improve the reliability of rumor queries.
- the foregoing obtaining unit 201, the query unit 202, the display unit 203, and the like may be embedded in or independent of the information query device in hardware, or may be stored in software.
- the processor is called to perform the operations corresponding to the above modules.
- the processor can be a central processing unit (CPU), a microprocessor, a microcontroller, or the like.
- FIG. 3 is a schematic block diagram showing a terminal provided by a fifth embodiment of the present invention.
- the terminal as shown may include: one or more processors 301 (only one shown); one or more input devices 302 (only one shown), one or more output devices 303 ( Only one) memory 304 is shown.
- the above processor 301, input device 302, output device 303, and memory 304 are connected by a bus 305.
- Memory 304 is used to store instructions
- processor 301 is used to execute instructions stored in memory 304. among them:
- the processor 301 is configured to: obtain, by the input device 302, the information to be queried; perform a matching query based on the to-be-queried information in a preset rumor database, where the rumor database includes: from two or more The plagiar information captured in the smashing website; when the matching phishing information is queried, the matching rumbling information is displayed by the output device 303.
- the processor 301 performs a matching query based on the to-be-queried information in the preset rumor database, including: using the M matching algorithms to separately query the to-be-queried information and the rumor database Performing a similarity calculation on each piece of plagiar information, obtaining M similarity scores of the information to be queried and the pieces of rumbling information, wherein the M is greater than or equal to 2 natural numbers;
- the weights set by the algorithm respectively weight and sum the M similarity scores of the to-be-queried information and the pieces of rumbling information, and obtain The total score of the similarity between the information to be inquired and the pieces of rumbling information is obtained; the plagiar information with the total score of the similarity higher than the preset threshold is determined as the matched rumbling information.
- the processor 301 is further configured to: when there is no similarity total score, the threshold is higher than the threshold The plagiar information ⁇ , sorting the pieces of placard information according to the order of the similarity total scores; displaying the first N plagiar information according to the sorted result, wherein the N is a preset positive integer.
- the rumor information is specifically a document title
- the rumor database further includes: a document link address associated with the rumbling information
- the processor 301 is further configured to: based on the preset rumor database After the query information is read and the matching query is performed, the document link address associated with the plagiar information to be displayed is displayed.
- the obtaining the to-be-queried information includes: receiving a rumor query instruction; and acquiring, by the rumor query instruction, the information selected by the current cursor as the information to be queried.
- the so-called processor 301 may be a central processing unit (CPU) and/or a graphics processing unit (GPU), or may be based thereon. Combined with other general-purpose processors, Digital Signal Processors (DSPs), Application Specific Integrated Circuits (ASICs), Field-Programmable Gate Arrays (FPGAs), or other programmable logic devices , discrete gates or transistor logic devices, discrete hardware components, etc.
- CPU central processing unit
- GPU graphics processing unit
- DSPs Digital Signal Processors
- ASICs Application Specific Integrated Circuits
- FPGAs Field-Programmable Gate Arrays
- programmable logic devices discrete gates or transistor logic devices, discrete hardware components, etc.
- the input device 302 may include a touch panel, a fingerprint sensor (for collecting fingerprint information of the user and direction information of the fingerprint), a microphone, a communication module (such as a Wi-Fi module, a 2G/3G/4G network module), Physical buttons, etc.
- a fingerprint sensor for collecting fingerprint information of the user and direction information of the fingerprint
- a microphone for collecting fingerprint information of the user and direction information of the fingerprint
- a communication module such as a Wi-Fi module, a 2G/3G/4G network module
- Physical buttons etc.
- the output device 303 may include a display (LCD or the like), a speaker, and the like.
- the display can be used to display information input by the user or information provided to the user, and the like.
- the display can include a display panel, optional, and can be used with a liquid crystal display (Liquid Crystal)
- Display panel is configured in the form of Display, LCD, or Organic Light-Emitting Diode (OLED). Further, the touch panel may be covered on the display. When the touch panel detects a touch operation on or near the touch panel, the touch panel transmits to the processor 301 to determine the type of the touch event, and then the processor 301 according to the type of the touch event. Provide the corresponding visual output on the display.
- the processor 301, the input device 302, the output device 303, and the memory 304 described in the embodiment of the present invention may be configured as described in the method embodiment of the information query method provided by the embodiment of the present invention. The implementation method will not be described here.
- the present invention further provides a storage medium, which may be a computer readable storage medium included in the memory in the above embodiment; or may be a computer readable storage that exists separately and is not assembled in the terminal. medium.
- the storage medium may be a non-transitory computer readable storage medium.
- the storage medium stores one or more programs, and the one or more programs are used by one or more processors to execute an information processing method, the method comprising: the embodiment further provides a computer readable storage medium
- the computer readable storage medium may be a computer readable storage medium included in the memory in the above embodiment; or may be a computer readable storage medium that is separately present and not incorporated in the terminal.
- the computer readable storage medium stores one or more programs, which may be executed by one or more processors for:
- the matching query is performed based on the to-be-queried information in a preset rumor database, including
- the rumbling information having the similarity total score higher than the preset threshold is determined as the matching rumbling information.
- the one or more programs may also be processed by the one or more processes. Execution to use In:
- the plagiar information is sorted according to the order of the similarity total scores from high to low;
- the rumor information is specifically a document title
- the rumor database further includes: a document link address associated with the rumbling information;
- the one or more programs may also be executed by the one or more processors for: displaying The link address of the document associated with the rumbling information to be displayed.
- the obtaining the information to be queried includes:
- the information selected by the current cursor is obtained as the information to be queried.
- the disclosed apparatus and method may be implemented in other manners.
- the device embodiments described above are merely illustrative.
- the division of the above units is only a logical function division, and the actual implementation may have another division manner, for example, multiple units or components may be combined or may be Integration into another system, or some features can be ignored, or not executed.
- the mutual coupling or direct coupling or communication connection shown or discussed may be an indirect coupling or communication connection through some interface, device or unit, and may be electrical, mechanical or otherwise.
Landscapes
- Engineering & Computer Science (AREA)
- Theoretical Computer Science (AREA)
- Data Mining & Analysis (AREA)
- Databases & Information Systems (AREA)
- Physics & Mathematics (AREA)
- General Engineering & Computer Science (AREA)
- General Physics & Mathematics (AREA)
- Information Retrieval, Db Structures And Fs Structures Therefor (AREA)
- User Interface Of Digital Computer (AREA)
Abstract
一种信息查询方法、信息查询装置、存储介质及终端,其中,上述信息查询方法包括:获取待查询信息;在预设的谣言数据库中基于所述待查询信息进行匹配查询,其中,所述谣言数据库包含:从两个以上辟谣网站中抓取的辟谣信息;当查询到匹配的辟谣信息时,显示所述匹配的辟谣信息。本技术方案能够有效提高谣言查询的可靠性。
Description
发明名称:信息査询方法、 信息査询装置、 存储介质及终端 技术领域
[0001] 本发明涉及信息技术领域, 具体涉及一种信息査询方法、 信息査询装置、 存储 介质及终端。
背景技术
[0002] 互联网的快速发展加快了信息的传播速度, 使得用户足不出户就可以了解到各 种动态, 然而, 由于我国的各种网络相关的管理措施或法律法规的不成熟, 在 信息快速传统的同吋, 很多不法分子利用网络热点事件编撰或凭空捏造一些谣 言以达到其不法的目的。
[0003] 为了防止更多的用户被欺骗, 以辟谣网站的形式出现的谣言引擎孕育而生, 当 用户需要査询某一信息是否为谣言吋, 用户可以在辟谣网站中输入该信息的关 键字来査询该信息的真实性。
[0004] 然而, 由于单一网站的信息局限性, 通过单一网站难以为用户提供全面的辟谣 信息, 这使得以辟谣网站的形式出现的谣言弓 I擎的辟谣可靠性较差。
技术问题
[0005] 本发明提供一种信息査询方法、 信息査询装置、 存储介质及终端, 用于提高谣 言査询的可靠性。
问题的解决方案
技术解决方案
[0006] 本发明一方面提供一种信息査询方法, 包括:
[0007] 获取待査询信息;
[0008] 在预设的谣言数据库中基于所述待査询信息进行匹配査询, 其中, 所述谣言数 据库包含: 从两个以上辟谣网站中抓取的辟谣信息;
[0009] 当査询到匹配的辟谣信息吋, 显示所述匹配的辟谣信息。
[0010] 基于上述第一方面, 在第一种可能的实现方式中, 所述在预设的谣言数据库中 基于所述待査询信息进行匹配査询, 包括:
[0011] 采用 M种匹配算法分别对所述待査询信息与所述谣言数据库中各条辟谣信息进 行相似度计算, 获得所述待査询信息与所述各条辟谣信息的 M个相似度分值, 其 中, 所述 M大于或等于 2的自然数;
[0012] 根据为所述各种匹配算法设定的权值分别对所述待査询信息与所述各条辟谣信 息的 M个相似度分值进行加权求和, 获得所述待査询信息与所述各条辟谣信息的 相似度总分值;
[0013] 将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣信息。
[0014] 基于上述第一方面的第一种可能的实现方式, 在第二种可能的实现方式中, 所 述获得所述待査询信息与所述各条辟谣信息的相似度总分值, 之后还包括: [0015] 当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似度总分值由高到 低的顺序对所述各条辟谣信息进行排序;
[0016] 根据排序的结果显示前 N条辟谣信息, 其中, 所述 N为预设的正整数。
[0017] 基于上述第一方面, 或者上述第一方面的第一种可能的实现方式, 或者上述第 一方面的第二种可能的实现方式, 在第三种可能的实现方式中, 所述辟谣信息 具体为文档标题, 所述谣言数据库还包含: 与所述辟谣信息关联的文档链接地 址;
[0018] 所述在预设的谣言数据库中基于所述待査询信息进行匹配査询, 之后还包括: 显示与需显示的辟谣信息关联的文档链接地址。
[0019] 基于上述第一方面, 或者上述第一方面的第一种可能的实现方式, 或者上述第 一方面的第二种可能的实现方式, 在第四种可能的实现方式中, 所述获取待査 询信息包括:
[0020] 接收谣言査询指令;
[0021] 在所述谣言査询指令的触发下, 获取当前光标所选定的信息作为待査询信息。
[0022] 本发明第二方面提供一种信息査询装置, 包括:
[0023] 获取单元, 用于获取待査询信息;
[0024] 査询单元, 用于在预设的谣言数据库中基于所述待査询信息进行匹配査询, 其 中, 所述谣言数据库包含: 从两个以上辟谣网站中抓取的辟谣信息;
[0025] 显示单元, 用于当所述査询单元査询到匹配的辟谣信息吋, 显示所述匹配的辟
谣信息。
[0026] 基于本发明第二方面, 在第一种可能的实现方式中, 所述査询单元包括: [0027] 相似度分值计算单元, 用于采用 M种匹配算法分别对所述待査询信息与所述谣 言数据库中各条辟谣信息进行相似度计算, 获得所述待査询信息与所述各条辟 谣信息的 M个相似度分值, 其中, 所述 M大于或等于 2的自然数;
[0028] 加权求和单元, 用于根据为所述各种匹配算法设定的权值分别对所述待査询信 息与所述各条辟谣信息的 M个相似度分值进行加权求和, 获得所述待査询信息与 所述各条辟谣信息的相似度总分值;
[0029] 确定单元, 用于将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣 f π息。
[0030] 基于本发明第二方面的第一种可能的实现方式, 在第二种可能的实现方式中, 所述信息査询装置还包括:
[0031] 排序单元, 用于当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似 度总分值由高到低的顺序对所述各条辟谣信息进行排序;
[0032] 所述显示单元还用于: 根据所述排序单元排序的结果显示前 N条辟谣信息, 其 中, 所述 N为预设的正整数。
[0033] 基于本发明第二方面, 或者本发明第二方面的第一种可能的实现方式, 或者本 发明第二方面的第二种可能的实现方式, 在第三种可能的实现方式中, 所述辟 谣信息具体为文档标题, 所述谣言数据库还包含: 与所述辟谣信息关联的文档 链接地址;
[0034] 所述显示单元还用于: 显示与需显示的辟谣信息关联的文档链接地址。
[0035] 基于本发明第二方面, 或者本发明第二方面的第一种可能的实现方式, 或者本 发明第二方面的第二种可能的实现方式, 在第四种可能的实现方式中, 所述获 取单元包括:
[0036] 指令接收单元, 用于接收谣言査询指令;
[0037] 子获取单元, 用于在所述指令接收单元接收到的所述谣言査询指令的触发下, 获取当前光标所选定的信息作为待査询信息。
[0038] 本发明第三方面提供一种存储介质, 所述存储介质存储有一个或者一个以上程
序, 所述一个或者一个以上程序可被一个或者一个以上的处理器执行以用于:
[0039] 获取待査询信息;
[0040] 在预设的谣言数据库中基于所述待査询信息进行匹配査询, 其中, 所述谣言数 据库包含: 从两个以上辟谣网站中抓取的辟谣信息;
[0041] 当査询到匹配的辟谣信息吋, 显示所述匹配的辟谣信息。
[0042] 基于上述第三方面, 在第一种可能的实现方式中, 所述在预设的谣言数据库中 基于所述待査询信息进行匹配査询, 包括:
[0043] 采用 M种匹配算法分别对所述待査询信息与所述谣言数据库中各条辟谣信息进 行相似度计算, 获得所述待査询信息与所述各条辟谣信息的 M个相似度分值, 其 中, 所述 M大于或等于 2的自然数;
[0044] 根据为所述各种匹配算法设定的权值分别对所述待査询信息与所述各条辟谣信 息的 M个相似度分值进行加权求和, 获得所述待査询信息与所述各条辟谣信息的 相似度总分值;
[0045] 将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣信息。
[0046] 基于上述第三方面的第一种可能的实现方式, 在第二种可能的实现方式中, 在 所述获得所述待査询信息与所述各条辟谣信息的相似度总分值之后, 所述一个 或者一个以上程序还可被所述一个或者一个以上的处理器执行以用于:
[0047] 当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似度总分值由高到 低的顺序对所述各条辟谣信息进行排序;
[0048] 根据排序的结果显示前 N条辟谣信息, 其中, 所述 N为预设的正整数。
[0049] 基于上述第三方面, 或者上述第三方面的第一种可能的实现方式, 或者上述第 三方面的第二种可能的实现方式, 在第三种可能的实现方式中, 所述辟谣信息 具体为文档标题, 所述谣言数据库还包含: 与所述辟谣信息关联的文档链接地 址;
[0050] 所述在预设的谣言数据库中基于所述待査询信息进行匹配査询之后, 所述一个 或者一个以上程序还可被所述一个或者一个以上的处理器执行以用于: 显示与 需显示的辟谣信息关联的文档链接地址。
[0051] 基于上述第三方面, 或者上述第三方面的第一种可能的实现方式, 或者上述第
三方面的第二种可能的实现方式, 在第四种可能的实现方式中, 所述获取待査 询信息包括:
[0052] 接收谣言査询指令;
[0053] 在所述谣言査询指令的触发下, 获取当前光标所选定的信息作为待査询信息。
[0054] 本发明第四方面提供一种终端, 所述终端包括: 至少一个处理器, 至少一个输 入设备, 至少一个输出设备, 以及存储器;
[0055] 所述存储器用于存储指令, 所述处理器用于执行所述存储器存储的指令; 其中
, 所述处理器用于:
[0056] 通过所述输入设备获取待査询信息;
[0057] 在预设的谣言数据库中基于所述待査询信息进行匹配査询, 其中, 所述谣言数 据库包含: 从两个以上辟谣网站中抓取的辟谣信息;
[0058] 当査询到匹配的辟谣信息吋, 通过所述输出设备显示所述匹配的辟谣信息。
[0059] 基于上述第四方面, 在第一种可能的实现方式中, 所述在预设的谣言数据库中 基于所述待査询信息进行匹配査询, 包括:
[0060] 采用 M种匹配算法分别对所述待査询信息与所述谣言数据库中各条辟谣信息进 行相似度计算, 获得所述待査询信息与所述各条辟谣信息的 M个相似度分值, 其 中, 所述 M大于或等于 2的自然数;
[0061] 根据为所述各种匹配算法设定的权值分别对所述待査询信息与所述各条辟谣信 息的 M个相似度分值进行加权求和, 获得所述待査询信息与所述各条辟谣信息的 相似度总分值;
[0062] 将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣信息。
[0063] 基于上述第四方面的第一种可能的实现方式, 在第二种可能的实现方式中, 在 获得所述待査询信息与所述各条辟谣信息的相似度总分值之后, 所述处理器还 用于:
[0064] 当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似度总分值由高到 低的顺序对所述各条辟谣信息进行排序;
[0065] 根据排序的结果显示前 N条辟谣信息, 其中, 所述 N为预设的正整数。
[0066] 基于上述第四方面, 或者上述第四方面的第一种可能的实现方式, 或者上述第
四方面的第二种可能的实现方式, 在第三种可能的实现方式中, 所述辟谣信息 为文档标题, 所述谣言数据库还包含: 与所述辟谣信息关联的文档链接地址;
[0067] 所述处理器还用于: 在预设的谣言数据库中基于所述待査询信息进行匹配査询 之后, 显示与需显示的辟谣信息关联的文档链接地址。
[0068] 基于上述第四方面, 或者上述第四方面的第一种可能的实现方式, 或者上述第 四方面的第二种可能的实现方式, 在第四种可能的实现方式中, 所述获取待査 询信息包括:
[0069] 接收谣言査询指令;
[0070] 在所述谣言査询指令的触发下, 获取当前光标所选定的信息作为待査询信息。
发明的有益效果
有益效果
[0071] 由上可见, 本发明中当获取到待査询信息吋, 在预设的谣言数据库中基于所述 待査询信息进行匹配査询并在査询到匹配的辟谣信息吋显示该匹配的辟谣信息 , 由于该谣言数据库是通过从多个辟谣网站中抓取的辟谣信息构成的大数据的 谣言数据库, 因此, 相对于传统的谣言引擎, 本发明能够提供更全面的辟谣信 息, 从而有效提高了谣言査询的可靠性。
对附图的简要说明
附图说明
[0072] 为了更清楚地说明本发明实施例或现有技术中的技术方案, 下面将对实施例或 现有技术描述中所需要使用的附图作简单地介绍, 显而易见地, 下面描述中的 附图仅仅是本发明的一些实施例, 对于本领域普通技术人员来讲, 在不付出创 造性劳动性的前提下, 还可以根据这些附图获得其他的附图。
[0073] 图 1-a为本发明提供的一种信息査询方法一个实施例流程示意图;
[0074] 图 1-b为本发明提供的一种待査询信息输入的界面示意图;
[0075] 图 1-c为本发明提供的另一种待査询信息输入的界面示意图;
[0076] 图 2为本发明提供的一种信息査询装置一个实施例结构示意图;
[0077] 图 3为本发明提供的一种终端一个实施例结构示意图。
本发明的实施方式
[0078] 为使得本发明的发明目的、 特征、 优点能够更加的明显和易懂, 下面将结合本 发明实施例中的附图, 对本发明实施例中的技术方案进行清楚、 完整地描述, 显然, 所描述的实施例仅仅是本发明一部分实施例, 而非全部实施例。 基于本 发明中的实施例, 本领域普通技术人员在没有做出创造性劳动前提下所获得的 所有其他实施例, 都属于本发明保护的范围。
[0079] 实施例一
[0080] 本发明实施例提供一种信息査询方法, 请参阅图 l-a, 本发明实施例中的信息 査询方法, 包括:
[0081] 步骤 101、 获取待査询信息;
[0082] 本发明实施例中, 待査询信息可以由一个或两个以上关键词构成, 或者也可以 由字符串构成, 此处不作限定。
[0083] 在一种应用场景中, 为用户提供文本输入控件, 以便用户通过该文本输入控件 输入待査询信息, 当接收到触发图 1-a所示的信息査询方法的指令吋, 获取该文 本输入控件中输入的信息作为待査询信息。 举例说明, 如图 1-b所示的谣言査询 操作界面, 用户可以通过文本输入控件 11输入待査询信息 (例如 "XX地区发生枪 击事件") , 之后点击名为"査询"的按键 12触发图 1-a所示的信息査询方法, 步骤 101获取当前文本输入控件 11输入的 "XX地区发生枪击事件"作为待査询信息。
[0084] 在另一种应用场景中, 用户可以通过光标选定待査询信息, 并在选定之后输入 谣言査询指令, 以便通过该谣言査询指令触发步骤 101的执行, 则步骤 101具体 包括: 接收谣言査询指令, 在上述谣言査询指令的触发下, 获取当前光标所选 定的信息作为待査询信息。 举例说明, 如图 1-c所示, 用户可以通过光标选定待 査询信息 (例如加显底部分: "XX地区发生枪击事件") , 之后通过预设的方式
(例如长按被选定的区域) 激活菜单 21, 从菜单 21中点击名为"査询"的选项来输 入上述谣言査询指令。
[0085] 当然, 本发明实施例中也可以采用其它方式获取待査询信息, 此处不作限定。
[0086] 步骤 102、 在预设的谣言数据库中基于上述待査询信息进行匹配査询;
[0087] 其中, 上述谣言数据库包含: 从两个以上辟谣网站中抓取的辟谣信息。
[0088] 本发明实施例中, 从多个辟谣网站中抓取信息构成大数据的谣言数据库, 并基 于该谣言数据库形成谣言引擎。 具体地, 上述辟谣信息可以为文档标题, 则从 多个辟谣网站中抓取的信息可以以文档标题和对应的文档链接地址为主, 当然 , 上述辟谣信息也可以为具体的文档内容, 此处不作限定。
[0089] 在一种应用场景中, 步骤 102可采用汉语言解析式进行匹配査询, 步骤 102包括 : 采用 M种匹配算法分别对上述待査询信息与上述谣言数据库中各条辟谣信息进 行相似度计算, 获得上述待査询信息与上述各条辟谣信息的 M个相似度分值; 根 据为上述各种匹配算法设定的权值分别对上述待査询信息与上述各条辟谣信息 的 M个相似度分值进行加权求和, 获得上述待査询信息与上述各条辟谣信息的相 似度总分值; 将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣信 息。 其中, 上述 M大于或等于 2的自然数。 举例说明, 设上述谣言数据库中存在 第一辟谣信息和第二辟谣信息 (当然, 实际的谣言数据库包含的辟谣信息远不 止两条) , 为第一匹配算法设定权值 Sl, 为第二匹配算法设定权值 S2, 采用第 一匹配算法和第二匹配算法 (即 M取 2) 分别对上述待査询信息与上述谣言数据 库中各条辟谣信息进行相似度计算, 则可获得上述待査询信息与上述第一条辟 谣信息的两个相似度分值 (记为 al和 a2) , 以及上述待査询信息与上述第二条辟 谣信息的两个相似度分值 (记为 bl和 b2) , 根据为上述第一匹配算法和上述第 二匹配算法设定的权值分别对上述待査询信息与各条辟谣信息的两个相似度分 值进行加权求和, 获得上述待査询信息与上述第一条辟谣信息的相似度总分值 A 1 (其中, Al=al*Sl+a2*S2) , 以及上述待査询信息与上述第二条辟谣信息的相 似度总分值 A2 (其中, A2=bl*Sl+b2*S2) 。 假定上述阈值为 A', 且 Al< A'< A2 , 则在步骤 102中, 将相似度总分值高于 A'的辟谣信息确定为匹配的辟谣信息, 即, 将上述第二条辟谣信息确定为匹配的辟谣信息。
[0090] 其中, 上述 M种匹配算法可包括但不限于如下算法: 同义词编码算法、 最长子 串重合算法和最小编辑距离算法。 以下分别对这三种算法进行说明:
[0091] 同义词编码算法是一种基于分级构造的同义词典的算法, 该同义词典中, 越相 近的词联系越紧密, 编码距离也越小, 相似度分值也越大。 当得到待査询信息
之后进行分词, 在同义词典中査找该词和比较词。 得到相应的编码并进行距离 的计算, 即可得到基于语义的近义词分析。 举例说明, 存在 A词和 B词, A词在 该同义词典中的编码为 E111234, B词在该同义词典中的编码为 E111237 , 则 A-B 的编码距离 = 3; 若存在 A词和 C词, A词在该同义词典中的编码为 E111234, B词 在该同义词典中的编码为 E134234, 贝 IjA-C的编码距离 =23000。 B词与 A词的相似 度分值要大于 C词与 A词的相似度分值。
[0092] 最长子串重合算法对于要匹配的句子, 分析其字串中重复的词的个数, 以此得 到基于句法的相似度分析, 重复的词的个数越多, 相似度分值也越大。 举例说 明: A句子: 今天天气很好, B句子: 今天明天和后天, C句子: 今天天气很好 啊。 A句子和 B句子重复的词的个数为 2 (即"今天") ; A句子和 C句子重复的词 的个数为 6 (即"今天天气很好") , 则 C句子与 A句子的相似度分值要大于 B句子 与 A句子的相似度分值。
[0093] 最小编辑距离算法根据输入的句子, 分析将目标句子转化为输入的句子所需要 的最少修改次数, 按照最少修改次数进行相似度计算, 最少修改次数越多, 相 似度分值越小。 举例说明, A句子: 春天来了; B句子: 春天来了吗; C句子: 问题来了。 将 B句子修改为 A句子的最少修改次数为 1 (即, 把 B句子中的"吗"去 掉) , 将 C句子修改为 A句子的最少修改次数为 2 (即, 把 C句子中的"问"改成"春 "且把 C句子中的"题"改为"天") 。 贝 IjC句子与 A句子的相似度分值要小于 B句子与 A句子的相似度分值。
[0094] 当然, 本发明实施例中也可以采用其它方式在预设的谣言数据库中基于上述待 査询信息进行匹配査询, 例如, 采用预设的一匹配算法分别对上述待査询信息 与上述谣言数据库中各条辟谣信息进行相似度计算, 获得上述待査询信息与上 述各条辟谣信息的相似度分值, 将相似度分值高于预设的另一阈值的辟谣信息 确定为匹配的辟谣信息; 或者, 也可以采用搜素引擎式在上述谣言数据库中基 于上述待査询信息进行匹配査询, 则将上述谣言数据库索引成为搜索区, 利用 待査询信息进行主动的搜寻并获得相应的相似度分值, 对于相似度分值高于一 定分数的辟谣信息, 即确认为匹配的辟谣信息。 进一步, 当不存在相似度分值 高于一定分数的辟谣信息吋, 也可显示最为相近的几条谣言以备査询。
[0095] 步骤 103、 当査询到匹配的辟谣信息吋, 显示上述匹配的辟谣信息;
[0096] 本发明实施例中, 当査询到匹配的辟谣信息吋, 显示匹配到的辟谣信息, 以便 用户査阅。
[0097] 为了节约存储空间, 可选的, 上述辟谣信息具体为文档标题, 上述谣言数据库 还包含: 与上述辟谣信息关联的文档链接地址。 则在步骤 102之后, 可进一步显 示与需显示的辟谣信息关联的文档链接地址, 以便用户通过访问显示的文档链 接地址査阅具体的文档内容。
[0098] 可选的, 当步骤 102中采用汉语言解析式进行匹配査询吋, 若査询到不存在相 似度总分值高于上述阈值的辟谣信息吋, 则按照相似度总分值由高到低的顺序 对上述各条辟谣信息进行排序, 并根据排序的结果显示前 N条辟谣信息, 其中, 上述 N为预设的正整数。
[0099] 需要说明的是, 本发明实施例中的信息査询方法具体可以由信息査询装置实现 , 该信息査询装置可以以软件形式 (例如 APP) 集成在智能终端中。 上述智能终 端具体可以为智能手机、 平板电脑、 个人计算机或其它电子终端, 此处不作限 定。
[0100] 由上可见, 本发明中当获取到待査询信息吋, 在预设的谣言数据库中基于所述 待査询信息进行匹配査询并在査询到匹配的辟谣信息吋显示该匹配的辟谣信息 , 由于该谣言数据库是通过从多个辟谣网站中抓取的辟谣信息构成的大数据的 谣言数据库, 因此, 相对于传统的谣言引擎, 本发明能够提供更全面的辟谣信 息, 从而有效提高了谣言査询的可靠性。
[0101] 需要说明的是, 本领域普通技术人员可以理解实现上述实施例的全部或部分步 骤可以通过硬件来完成, 也可以通过程序来指令相关的硬件完成, 所述的程序 可以存储于一种计算机可读存储介质中, 上述提到的存储介质可以是只读存储 器, 磁盘或光盘等。
[0102] 实施例二
[0103] 本发明实施例提供一种信息査询装置。 请参阅图 2, 本发明实施例中的信息査 询装置 200, 包括:
[0104] 获取单元 201, 用于获取待査询信息;
[0105] 査询单元 202, 用于在预设的谣言数据库中基于所述待査询信息进行匹配査询 , 其中, 所述谣言数据库包含: 从两个以上辟谣网站中抓取的辟谣信息;
[0106] 显示单元 203, 用于当査询单元 202査询到匹配的辟谣信息吋, 显示所述匹配的 辟谣信息。
[0107] 可选的, 査询单元 202包括:
[0108] 相似度分值计算单元, 用于采用 M种匹配算法分别对所述待査询信息与所述谣 言数据库中各条辟谣信息进行相似度计算, 获得所述待査询信息与所述各条辟 谣信息的 M个相似度分值, 其中, 所述 M大于或等于 2的自然数;
[0109] 加权求和单元, 用于根据为所述各种匹配算法设定的权值分别对所述待査询信 息与所述各条辟谣信息的 M个相似度分值进行加权求和, 获得所述待査询信息与 所述各条辟谣信息的相似度总分值;
[0110] 确定单元, 用于将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣 f π息。
[0111] 进一步, 本发明实施例中的信息査询装置还包括: 排序单元, 用于当不存在相 似度总分值高于所述阈值的辟谣信息吋, 按照相似度总分值由高到低的顺序对 所述各条辟谣信息进行排序; 所述显示单元还用于: 根据所述排序单元排序的 结果显示前 N条辟谣信息, 其中, 所述 N为预设的正整数。
[0112] 可选的, 所述辟谣信息具体为文档标题, 所述谣言数据库还包含: 与所述辟谣 信息关联的文档链接地址; 显示单元 203还用于: 显示与需显示的辟谣信息关联 的文档链接地址。
[0113] 可选的, 获取单元 201包括:
[0114] 指令接收单元, 用于接收谣言査询指令;
[0115] 子获取单元, 用于在所述指令接收单元接收到的所述谣言査询指令的触发下, 获取当前光标所选定的信息作为待査询信息。
[0116] 需要说明的是, 本发明实施例中的信息査询装置可以以软件形式 (例如 APP) 集成在智能终端中。 上述智能终端具体可以为智能手机、 平板电脑、 个人计算 机或其它电子终端, 此处不作限定。
[0117] 应理解, 本发明实施例中的信息査询装置的各个功能模块的功能可以根据上述
方法实施例中的方法具体实现, 其具体实现过程可参照上述方法实施例中的相 关描述, 此处不再赘述。
[0118] 由上可见, 本发明中当获取到待査询信息吋, 在预设的谣言数据库中基于所述 待査询信息进行匹配査询并在査询到匹配的辟谣信息吋显示该匹配的辟谣信息 , 由于该谣言数据库是通过从多个辟谣网站中抓取的辟谣信息构成的大数据的 谣言数据库, 因此, 相对于传统的谣言引擎, 本发明能够提供更全面的辟谣信 息, 从而有效提高了谣言査询的可靠性。
[0119] 需要说明的是, 在硬件实现上, 以上获取单元 201, 査询单元 202, 显示单元 20 3等可以以硬件形式内嵌于或独立于信息査询装置中, 也可以以软件形式存储于 信息査询装置的存储器中, 以便于处理器调用执行以上各个模块对应的操作。 该处理器可以为中央处理单元 (CPU) 、 微处理器、 单片机等。
[0120] 实施例三
[0121] 为了便于更好地实施本发明实施例中的上述方法实施例, 本发明还提供了用于 配合实施执行上述方法实施例的相关终端。 图 3给出本发明第五实施例提供的终 端的示意性框图。 如图所示的该终端可以包括: 一个或多个处理器 301 (图中仅 示出一个) ; 一个或多个输入设备 302 (图中仅示出一个) , 一个或多个输出设 备 303 (图中仅示出一个) 、 存储器 304。 上述处理器 301、 输入设备 302、 输出 设备 303、 存储器 304通过总线 305连接。 存储器 304用于存储指令, 处理器 301用 于执行存储器 304存储的指令。 其中:
[0122] 处理器 301用于: 通过输入设备 302获取待査询信息; 在预设的谣言数据库中基 于所述待査询信息进行匹配査询, 其中, 所述谣言数据库包含: 从两个以上辟 谣网站中抓取的辟谣信息; 当査询到匹配的辟谣信息吋, 通过输出设备 303显示 所述匹配的辟谣信息。
[0123] 可选的, 处理器 301在预设的谣言数据库中基于所述待査询信息进行匹配査询 , 包括: 采用 M种匹配算法分别对所述待査询信息与所述谣言数据库中各条辟谣 信息进行相似度计算, 获得所述待査询信息与所述各条辟谣信息的 M个相似度分 值, 其中, 所述 M大于或等于 2的自然数; 根据为所述各种匹配算法设定的权值 分别对所述待査询信息与所述各条辟谣信息的 M个相似度分值进行加权求和, 获
得所述待査询信息与所述各条辟谣信息的相似度总分值; 将相似度总分值高于 预设的阈值的辟谣信息确定为匹配的辟谣信息。
[0124] 可选的, 在获得所述待査询信息与所述各条辟谣信息的相似度总分值之后, 处 理器 301还用于: 当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似 度总分值由高到低的顺序对所述各条辟谣信息进行排序; 根据排序的结果显示 前 N条辟谣信息, 其中, 所述 N为预设的正整数。
[0125] 可选的, 所述辟谣信息具体为文档标题, 所述谣言数据库还包含: 与所述辟谣 信息关联的文档链接地址; 处理器 301还用于: 在预设的谣言数据库中基于所述 待査询信息进行匹配査询之后, 显示与需显示的辟谣信息关联的文档链接地址
[0126] 可选的, 所述获取待査询信息包括: 接收谣言査询指令; 在所述谣言査询指令 的触发下, 获取当前光标所选定的信息作为待査询信息。
[0127] 应当理解, 在本发明实施例中, 所称处理器 301可以是中央处理单元 (Central Processing Unit, CPU)和 /或图形处理器 (Graphic Processing Unit, GPU) , 也 可以在此基础上结合其他通用处理器、 数字信号处理器(Digital Signal Processor , DSP)、 专用集成电路(Application Specific Integrated Circuit, ASIC)、 现成可编 程门阵列(Field-Programmable Gate Array , FPGA)或者其他可编程逻辑器件、 分 立门或者晶体管逻辑器件、 分立硬件组件等。
[0128] 输入设备 302可以包括触控板、 指纹采传感器 (用于采集用户的指纹信息和指 纹的方向信息) 、 麦克风、 通信模块 (比如 Wi-Fi模块、 2G/3G/4G网络模块) 、 物理按键等。
[0129] 输出设备 303可以包括显示器 (LCD等) 、 扬声器等。 其中, 显示器可用于显 示由用户输入的信息或提供给用户的信息等。 显示器可包括显示面板, 可选的 , 可以采用液晶显示器 (Liquid Crystal
Display , LCD) 、 有机发光二极管 (Organic Light-Emitting Diode, OLED) 等形 式来配置显示面板。 进一步的, 上述触控板可覆盖在显示器上, 当触控板检测 到在其上或附近的触摸操作后, 传送给处理器 301以确定触摸事件的类型, 随后 处理器 301根据触摸事件的类型在显示器上提供相应的视觉输出。
[0130] 具体实现中, 本发明实施例中所描述的处理器 301、 输入设备 302、 输出设备 30 3、 存储器 304可执行本发明实施例提供的信息査询方法的方法实施例中所描述 的实现方式, 在此不再赘述。
[0131] 实施例四
[0132] 本发明还提供了一种存储介质, 该存储介质可以是上述实施例中的存储器中所 包含的计算机可读存储介质; 也可以是单独存在, 未装配入终端中的计算机可 读存储介质。 具体地, 该存储介质可以为非暂态计算机可读存储介质。 上述存 储介质存储有一个或者一个以上程序, 所述一个或者一个以上程序被一个或者 一个以上的处理器用来执行一个信息处理方法, 所述方法包括: 实施例还提供 了一种计算机可读存储介质, 该计算机可读存储介质可以是上述实施例中的存 储器中所包含的计算机可读存储介质; 也可以是单独存在, 未装配入终端中的 计算机可读存储介质。 所述计算机可读存储介质存储有一个或者一个以上程序 , 所述一个或者一个以上程序可被一个或者一个以上的处理器执行以用于:
[0133] 获取待査询信息;
[0134] 在预设的谣言数据库中基于所述待査询信息进行匹配査询, 其中, 所述谣言数 据库包含: 从两个以上辟谣网站中抓取的辟谣信息;
[0135] 当査询到匹配的辟谣信息吋, 显示所述匹配的辟谣信息。
[0136] 可选的, 所述在预设的谣言数据库中基于所述待査询信息进行匹配査询, 包括
[0137] 采用 M种匹配算法分别对所述待査询信息与所述谣言数据库中各条辟谣信息进 行相似度计算, 获得所述待査询信息与所述各条辟谣信息的 M个相似度分值, 其 中, 所述 M大于或等于 2的自然数;
[0138] 根据为所述各种匹配算法设定的权值分别对所述待査询信息与所述各条辟谣信 息的 M个相似度分值进行加权求和, 获得所述待査询信息与所述各条辟谣信息的 相似度总分值;
[0139] 将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣信息。
[0140] 可选的, 在所述获得所述待査询信息与所述各条辟谣信息的相似度总分值之后 , 所述一个或者一个以上程序还可被所述一个或者一个以上的处理器执行以用
于:
[0141] 当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似度总分值由高到 低的顺序对所述各条辟谣信息进行排序;
[0142] 根据排序的结果显示前 N条辟谣信息, 其中, 所述 N为预设的正整数。
[0143] 可选的, 所述辟谣信息具体为文档标题, 所述谣言数据库还包含: 与所述辟谣 信息关联的文档链接地址;
[0144] 所述在预设的谣言数据库中基于所述待査询信息进行匹配査询之后, 所述一个 或者一个以上程序还可被所述一个或者一个以上的处理器执行以用于: 显示与 需显示的辟谣信息关联的文档链接地址。
[0145] 可选的, 所述获取待査询信息包括:
[0146] 接收谣言査询指令;
[0147] 在所述谣言査询指令的触发下, 获取当前光标所选定的信息作为待査询信息。
[0148] 需要说明的是, 在本申请所提供的几个实施例中, 应该理解到, 所揭露的装置 和方法, 可以通过其它的方式实现。 例如, 以上所描述的装置实施例仅仅是示 意性的, 例如, 上述单元的划分, 仅仅为一种逻辑功能划分, 实际实现吋可以 有另外的划分方式, 例如多个单元或组件可以结合或者可以集成到另一个系统 , 或一些特征可以忽略, 或不执行。 另一点, 所显示或讨论的相互之间的耦合 或直接耦合或通信连接可以是通过一些接口, 装置或单元的间接耦合或通信连 接, 可以是电性, 机械或其它的形式。
[0149] 对于前述的各方法实施例, 为了简便描述, 故将其都表述为一系列的动作组合 , 但是本领域技术人员应该知悉, 本发明并不受所描述的动作顺序的限制, 因 为依据本发明, 某些步骤可以采用其它顺序或者同吋进行。 其次, 本领域技术 人员也应该知悉, 说明书中所描述的实施例均属于优选实施例, 所涉及的动作 和模块并不一定都是本发明所必须的。
[0150] 在上述实施例中, 对各个实施例的描述都各有侧重, 某个实施例中没有详述的 部分, 可以参见其它实施例的相关描述。
[0151] 以上为对本发明所提供的一种信息査询方法、 信息査询装置、 存储介质及终端 的描述, 对于本领域的一般技术人员, 依据本发明实施例的思想, 在具体实施
方式及应用范围上均会有改变之处, 综上, 本说明书内容不应理解为对本发明 的限制。
Claims
[权利要求 1] 一种信息査询方法, 其特征在于, 包括:
获取待査询信息;
在预设的谣言数据库中基于所述待査询信息进行匹配査询, 其中, 所 述谣言数据库包含: 从两个以上辟谣网站中抓取的辟谣信息; 当査询到匹配的辟谣信息吋, 显示所述匹配的辟谣信息。
[权利要求 2] 根据权利要求 1所述的方法, 其特征在于, 所述在预设的谣言数据库 中基于所述待査询信息进行匹配査询, 包括:
采用 M种匹配算法分别对所述待査询信息与所述谣言数据库中各条辟 谣信息进行相似度计算, 获得所述待査询信息与所述各条辟谣信息的 M个相似度分值, 其中, 所述 M大于或等于 2的自然数;
根据为所述各种匹配算法设定的权值分别对所述待査询信息与所述各 条辟谣信息的 M个相似度分值进行加权求和, 获得所述待査询信息与 所述各条辟谣信息的相似度总分值;
将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣信息。
[权利要求 3] 根据权利要求 2所述的方法, 其特征在于, 所述获得所述待査询信息 与所述各条辟谣信息的相似度总分值, 之后还包括:
当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似度总分 值由高到低的顺序对所述各条辟谣信息进行排序; 根据排序的结果显示前 N条辟谣信息, 其中, 所述 N为预设的正整数
[权利要求 4] 根据权利要求 1至 3任一项所述的方法, 其特征在于, 所述辟谣信息具 体为文档标题, 所述谣言数据库还包含: 与所述辟谣信息关联的文档 链接地址;
所述在预设的谣言数据库中基于所述待査询信息进行匹配査询, 之后 还包括: 显示与需显示的辟谣信息关联的文档链接地址。
[权利要求 5] 根据权利要求 1至 3任一项所述的方法, 其特征在于, 所述获取待査询 信息包括:
接收谣言査询指令;
在所述谣言査询指令的触发下, 获取当前光标所选定的信息作为待査 询信息。
[权利要求 6] —种信息査询装置, 其特征在于, 包括:
获取单元, 用于获取待査询信息;
査询单元, 用于在预设的谣言数据库中基于所述待査询信息进行匹配 査询, 其中, 所述谣言数据库包含: 从两个以上辟谣网站中抓取的辟 谣信息;
显示单元, 用于当所述査询单元査询到匹配的辟谣信息吋, 显示所述 匹配的辟谣信息。
[权利要求 7] 根据权利要求 6所述的信息査询装置, 其特征在于,
所述査询单元包括:
相似度分值计算单元, 用于采用 M种匹配算法分别对所述待査询信息 与所述谣言数据库中各条辟谣信息进行相似度计算, 获得所述待査询 信息与所述各条辟谣信息的 M个相似度分值, 其中, 所述 M大于或等 于 2的自然数;
加权求和单元, 用于根据为所述各种匹配算法设定的权值分别对所述 待査询信息与所述各条辟谣信息的 M个相似度分值进行加权求和, 获 得所述待査询信息与所述各条辟谣信息的相似度总分值;
确定单元, 用于将相似度总分值高于预设的阈值的辟谣信息确定为匹 配的辟谣信息。
[权利要求 8] 根据权利要求 7所述的信息査询装置, 其特征在于, 所述信息査询装 置还包括:
排序单元, 用于当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似度总分值由高到低的顺序对所述各条辟谣信息进行排序; 所述显示单元还用于: 根据所述排序单元排序的结果显示前 N条辟谣 信息, 其中, 所述 N为预设的正整数。
[权利要求 9] 根据权利要求 6至 8任一项所述的信息査询装置, 其特征在于, 所述辟
谣信息具体为文档标题, 所述谣言数据库还包含: 与所述辟谣信息关 联的文档链接地址;
所述显示单元还用于: 显示与需显示的辟谣信息关联的文档链接地址
[权利要求 10] 根据权利要求 6至 8任一项所述的信息査询装置, 其特征在于, 所述获 取单元包括:
指令接收单元, 用于接收谣言査询指令;
子获取单元, 用于在所述指令接收单元接收到的所述谣言査询指令的 触发下, 获取当前光标所选定的信息作为待査询信息。
[权利要求 11] 一种存储介质, 其特征在于, 所述存储介质存储有一个或者一个以上 程序, 所述一个或者一个以上程序可被一个或者一个以上的处理器执 行以用于:
获取待査询信息;
在预设的谣言数据库中基于所述待査询信息进行匹配査询, 其中, 所 述谣言数据库包含: 从两个以上辟谣网站中抓取的辟谣信息; 当査询到匹配的辟谣信息吋, 显示所述匹配的辟谣信息。
[权利要求 12] 根据权利要求 1所述的存储介质, 其特征在于, 所述在预设的谣言数 据库中基于所述待査询信息进行匹配査询, 包括: 采用 M种匹配算法分别对所述待査询信息与所述谣言数据库中各条辟 谣信息进行相似度计算, 获得所述待査询信息与所述各条辟谣信息的 M个相似度分值, 其中, 所述 M大于或等于 2的自然数;
根据为所述各种匹配算法设定的权值分别对所述待査询信息与所述各 条辟谣信息的 M个相似度分值进行加权求和, 获得所述待査询信息与 所述各条辟谣信息的相似度总分值;
将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣信息。
[权利要求 13] 根据权利要求 12所述的存储介质, 其特征在于, 在所述获得所述待査 询信息与所述各条辟谣信息的相似度总分值之后, 所述一个或者一个 以上程序还可被所述一个或者一个以上的处理器执行以用于:
当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似度总分 值由高到低的顺序对所述各条辟谣信息进行排序; 根据排序的结果显示前 N条辟谣信息, 其中, 所述 N为预设的正整数
[权利要求 14] 根据权利要求 11至 13任一项所述的存储介质, 其特征在于, 所述辟谣 信息为文档标题, 所述谣言数据库还包含: 与所述辟谣信息关联的文 档链接地址;
所述在预设的谣言数据库中基于所述待査询信息进行匹配査询之后, 所述一个或者一个以上程序还可被所述一个或者一个以上的处理器执 行以用于: 显示与需显示的辟谣信息关联的文档链接地址。
[权利要求 15] 根据权利要求 11至 13任一项所述的存储介质, 其特征在于, 所述获取 待査询信息包括:
接收谣言査询指令;
在所述谣言査询指令的触发下, 获取当前光标所选定的信息作为待査 询信息。
[权利要求 16] —种终端, 其特征在于, 所述终端包括: 至少一个处理器, 至少一个 输入设备, 至少一个输出设备, 以及存储器;
所述存储器用于存储指令, 所述处理器用于执行所述存储器存储的指 令; 其中, 所述处理器用于:
通过所述输入设备获取待査询信息;
在预设的谣言数据库中基于所述待査询信息进行匹配査询, 其中, 所 述谣言数据库包含: 从两个以上辟谣网站中抓取的辟谣信息; 当査询到匹配的辟谣信息吋, 通过所述输出设备显示所述匹配的辟谣 f π息。
[权利要求 17] 根据权利要求 16所述的终端, 其特征在于, 所述在预设的谣言数据库 中基于所述待査询信息进行匹配査询, 包括:
采用 M种匹配算法分别对所述待査询信息与所述谣言数据库中各条辟 谣信息进行相似度计算, 获得所述待査询信息与所述各条辟谣信息的
M个相似度分值, 其中, 所述 M大于或等于 2的自然数; 根据为所述各种匹配算法设定的权值分别对所述待査询信息与所述各 条辟谣信息的 M个相似度分值进行加权求和, 获得所述待査询信息与 所述各条辟谣信息的相似度总分值;
将相似度总分值高于预设的阈值的辟谣信息确定为匹配的辟谣信息。
[权利要求 18] 根据权利要求 17所述的终端, 其特征在于, 在获得所述待査询信息与 所述各条辟谣信息的相似度总分值之后, 所述处理器还用于: 当不存在相似度总分值高于所述阈值的辟谣信息吋, 按照相似度总分 值由高到低的顺序对所述各条辟谣信息进行排序; 根据排序的结果显示前 N条辟谣信息, 其中, 所述 N为预设的正整数
[权利要求 19] 根据权利要求 16至 18任一项所述的终端, 其特征在于, 所述辟谣信息 具体为文档标题, 所述谣言数据库还包含: 与所述辟谣信息关联的文 档链接地址;
所述处理器还用于: 在预设的谣言数据库中基于所述待査询信息进行 匹配査询之后, 显示与需显示的辟谣信息关联的文档链接地址。
[权利要求 20] 根据权利要求 16至 18任一项所述的终端, 其特征在于, 所述获取待査 询信息包括:
接收谣言査询指令;
在所述谣言査询指令的触发下, 获取当前光标所选定的信息作为待査 询信息。
Applications Claiming Priority (2)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| CN201610578606.2 | 2016-07-20 | ||
| CN201610578606.2A CN107644029A (zh) | 2016-07-20 | 2016-07-20 | 信息查询方法及信息查询装置 |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| WO2018014543A1 true WO2018014543A1 (zh) | 2018-01-25 |
Family
ID=60991725
Family Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| PCT/CN2017/073389 Ceased WO2018014543A1 (zh) | 2016-07-20 | 2017-02-13 | 信息查询方法、信息查询装置、存储介质及终端 |
Country Status (2)
| Country | Link |
|---|---|
| CN (1) | CN107644029A (zh) |
| WO (1) | WO2018014543A1 (zh) |
Families Citing this family (4)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN110928425A (zh) * | 2018-09-17 | 2020-03-27 | 北京搜狗科技发展有限公司 | 信息监控方法及装置 |
| CN109388696B (zh) * | 2018-09-30 | 2021-07-23 | 北京字节跳动网络技术有限公司 | 删除谣言文章的方法、装置、存储介质及电子设备 |
| CN111506794A (zh) * | 2020-04-17 | 2020-08-07 | 腾讯科技(武汉)有限公司 | 一种基于机器学习的谣言管理方法和装置 |
| CN113536760B (zh) * | 2021-07-06 | 2023-09-26 | 中国科学院计算技术研究所 | 引述句和辟谣模式句引导的“谣言-辟谣文章”匹配方法及系统 |
Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2010024184A1 (ja) * | 2008-08-26 | 2010-03-04 | 日本電気株式会社 | 風評情報検出システム、風評情報検出方法及びプログラム |
| CN102314645A (zh) * | 2011-09-26 | 2012-01-11 | 深圳市络道科技有限公司 | 一种地址匹配方法及匹配系统 |
| CN103902621A (zh) * | 2012-12-28 | 2014-07-02 | 深圳先进技术研究院 | 一种鉴定网络谣言的方法和装置 |
| CN104679739A (zh) * | 2013-11-27 | 2015-06-03 | 江苏华御信息技术有限公司 | 一种非真实信息传播控制方法 |
| CN105045857A (zh) * | 2015-07-09 | 2015-11-11 | 中国科学院计算技术研究所 | 一种社交网络谣言识别方法及系统 |
Family Cites Families (1)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| CN102637163A (zh) * | 2011-01-09 | 2012-08-15 | 华东师范大学 | 一种基于语义的多层次本体匹配的控制方法及系统 |
-
2016
- 2016-07-20 CN CN201610578606.2A patent/CN107644029A/zh active Pending
-
2017
- 2017-02-13 WO PCT/CN2017/073389 patent/WO2018014543A1/zh not_active Ceased
Patent Citations (5)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2010024184A1 (ja) * | 2008-08-26 | 2010-03-04 | 日本電気株式会社 | 風評情報検出システム、風評情報検出方法及びプログラム |
| CN102314645A (zh) * | 2011-09-26 | 2012-01-11 | 深圳市络道科技有限公司 | 一种地址匹配方法及匹配系统 |
| CN103902621A (zh) * | 2012-12-28 | 2014-07-02 | 深圳先进技术研究院 | 一种鉴定网络谣言的方法和装置 |
| CN104679739A (zh) * | 2013-11-27 | 2015-06-03 | 江苏华御信息技术有限公司 | 一种非真实信息传播控制方法 |
| CN105045857A (zh) * | 2015-07-09 | 2015-11-11 | 中国科学院计算技术研究所 | 一种社交网络谣言识别方法及系统 |
Non-Patent Citations (1)
| Title |
|---|
| JIAO, YANG: "Internet Rumors", CHINA MASTER'S THESES FULL-TEXT DATABASE, 15 February 2015 (2015-02-15), pages 40 - 42 * |
Also Published As
| Publication number | Publication date |
|---|---|
| CN107644029A (zh) | 2018-01-30 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| CN110647614B (zh) | 智能问答方法、装置、介质及电子设备 | |
| US10210243B2 (en) | Method and system for enhanced query term suggestion | |
| CN110472027B (zh) | 意图识别方法、设备及计算机可读存储介质 | |
| Hamidian et al. | Rumor identification and belief investigation on twitter | |
| US11481428B2 (en) | Bullet screen content processing method, application server, and user terminal | |
| JP6554685B2 (ja) | 検索結果を提供するための方法及び装置 | |
| US20210168098A1 (en) | Providing local service information in automated chatting | |
| CN114116997A (zh) | 知识问答方法、装置、电子设备及存储介质 | |
| US20130174058A1 (en) | System and Method to Automatically Aggregate and Extract Key Concepts Within a Conversation by Semantically Identifying Key Topics | |
| CN107832432A (zh) | 一种搜索结果排序方法、装置、服务器和存储介质 | |
| WO2020155747A1 (zh) | 问题答案推荐方法、装置、存储介质及服务器 | |
| WO2023168997A1 (zh) | 一种跨模态搜索方法及相关设备 | |
| WO2018058118A1 (en) | Method, apparatus and client of processing information recommendation | |
| CN107491465B (zh) | 用于搜索内容的方法和装置以及数据处理系统 | |
| EP3785144A1 (en) | Session message processing | |
| CN109710088B (zh) | 一种信息搜索方法及装置 | |
| WO2017206376A1 (zh) | 搜索方法、装置及非易失性计算机存储介质 | |
| WO2018014543A1 (zh) | 信息查询方法、信息查询装置、存储介质及终端 | |
| WO2020258481A1 (zh) | 个性化文本智能推荐方法、装置及计算机可读存储介质 | |
| CN111460783B (zh) | 一种数据处理方法、装置、计算机设备及存储介质 | |
| CN111324725B (zh) | 一种话题获取方法、终端、计算机可读存储介质 | |
| CN113094604B (zh) | 搜索结果排序方法、搜索方法及装置 | |
| CN108509059B (zh) | 一种信息处理方法、电子设备和计算机存储介质 | |
| CN113609372B (zh) | 搜索方法、装置、服务器、介质及产品 | |
| US20170161322A1 (en) | Method and electronic device for searching resource |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| 121 | Ep: the epo has been informed by wipo that ep was designated in this application |
Ref document number: 17830192 Country of ref document: EP Kind code of ref document: A1 |
|
| NENP | Non-entry into the national phase |
Ref country code: DE |
|
| 32PN | Ep: public notification in the ep bulletin as address of the adressee cannot be established |
Free format text: NOTING OF LOSS OF RIGHTS PURSUANT TO RULE 112(1) EPC (EPO FORM 1205A DATED 15.05.2019) |
|
| 122 | Ep: pct application non-entry in european phase |
Ref document number: 17830192 Country of ref document: EP Kind code of ref document: A1 |