Rumor detection system and method based on large language model

Through a rumor detection system based on a large language model, the core contradictions in the text are extracted and relevant credible judgment basis are retrieved, which solves the problem of inability to effectively utilize external objective facts and poor interpretability in the existing technology, and achieves more efficient and credible rumor detection.

CN120181073AActive Publication Date: 2025-06-20CHANGSHA ZHIWEI INFORMATION TECH CO LTD

Patent Information

Application Number
CN202510648140.8
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-05-20
Publication Date
2025-06-20
Estimated Expiration
2045-05-20

AI Technical Summary

Technical Problem

The existing rumor detection methods cannot effectively utilize external objective facts, and are poorly interpretable, making it difficult to distinguish core contradictions and interfered facts.

Method used

A rumor detection system based on a large language model is adopted, including a core contradiction extraction module, a fact verification and search module and a rumor judgment decision-making module. The core contradictions of entity relationships and attributes in the text to be detected are extracted through a large language model, and relevant credible judgment basis are retrieved using the knowledge base and the Internet, and finally, based on these basis, the detection results are formed.

Benefits of technology

It improves the accuracy of rumor detection and information utilization efficiency, reduces the probability of irrelevant information interference, and enhances the interpretability and persuasiveness of the detection results.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120181073A_ABST
    Figure CN120181073A_ABST
Patent Text Reader

Abstract

The invention discloses a rumor detection system based on a large language model.The rumor detection system comprises a core contradiction extraction module, a fact checking retrieval module and a rumor judgment decision module.The rumor detection system extracts core contradictions through the large language model, the probability of being interfered by irrelevant information can be reduced, follow-up retrieval is facilitated, the information utilization efficiency is improved, and the rumor detection efficiency is improved. The rumor detection precision is improved; meanwhile, related judgment bases are retrieved on a knowledge base and the internet through a fact checking retrieval module, sources of the judgment bases are enriched, meanwhile, it is avoided that illusion appears on a large language model to affect a detection result, and the reliability of the judgment bases is further improved through information source database filtering; and finally, the rumor judgment decision-making module enables the large language model to analyze the influence of each judgment basis abstract on the core contradiction, a detection result is formed, the human judgment process is simulated, and the detection result is based on the judgment basis and has interpretability.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present invention belongs to the technical field of information data processing, and particularly relates to a rumor detection system and method based on a large language model. Background Art

[0002] There is a vast amount of information on the Internet, and at the same time, the information evolves rapidly, making it inefficient for humans to judge and refute rumors. Therefore, automated rumor detection methods are needed to identify rumors from the vast amount of information on the Internet and the network environment.

[0003] Most of the early automated rumor identification technologies were based on text features and user metadata, including less data information, so the detection accuracy was insufficient. Subsequently, Bian et al. proposed a detection method based on the rumor propagation structure, which used a graph convolutional network to traverse from two directions of propagation, thereby capturing the features in the process of rumor text propagation and improving the rumor detection efficiency. Since Internet information is not limited to text data, but also includes various modalities such as pictures, audio, and video, Chen et al. separately abstracted multiple modalities such as text and images, and used the objective of reinforcement learning for optimization, improving the rumor detection ability for multi-modal information.

[0004] These rumor detection methods objectively improve the rumor detection efficiency, but their detection methods are all based on the information given by rumor fabricators, and these information often contain other facts used to confuse and interfere with the detection. Existing detection methods are difficult to effectively utilize external objective facts to distinguish the core contradictions from rumor information. At the same time, the existing rumor detection methods have poor interpretability of the detection results, reducing the persuasiveness of the detection results. Summary of the Invention

[0005] The technical problem to be solved by the present invention is to overcome the defects that the existing rumor detection methods cannot effectively utilize external objective facts and have poor interpretability, so as to provide a rumor detection system and method based on a large language model.

[0006] The present invention provides a rumor detection system based on a large language model, including: A core contradiction extraction module, configured to: input a first prompt word and a text to be detected into the large language model to extract the core contradiction of entity relationships in the text to be detected, where the core contradiction of entity relationships is the contradiction between the mutual relationships of multiple entities in the text to be detected output by the large language model and the facts; input a second prompt word and the text to be detected into the large language model to extract the core contradiction of entity attributes in the text to be detected, where the core contradiction of entity attributes is the contradiction between the attributes of the entity in the text to be detected output by the large language model and the facts; The fact-checking retrieval module includes a knowledge base retrieval module and an Internet retrieval module; the knowledge base retrieval module is used to: retrieve multiple pieces of knowledge base data most relevant to the core contradiction in the knowledge base through the RAG algorithm to obtain relevant knowledge base data; the Internet retrieval module is used to: retrieve the core contradiction from the Internet to obtain a retrieval result, filter the retrieval result through a source database to obtain a credible retrieval result; input the third prompt word, the core contradiction, and the credible retrieval result into a large language model to obtain relevant credible retrieval results; the core contradiction includes the entity relationship core contradiction and the entity attribute core contradiction; The rumor judgment and decision-making module is used to: input the fourth prompt word, the core contradiction, and the judgment basis into a large language model, summarize each judgment basis, and extract the content that mutually corroborates or contradicts the core contradiction therein to form a judgment basis summary; the judgment basis includes the relevant knowledge base data and the relevant credible retrieval results; input the fifth prompt word, the text to be detected, the core contradiction, and the judgment basis summary into a large language model, analyze the influence of each judgment basis summary on the core contradiction, and form a detection result.

[0007] Furthermore, the core contradiction extraction module is also used to: add its time mark when extracting the entity relationship core contradiction in the text to be detected, and add its time mark when extracting the entity attribute core contradiction in the text to be detected; The fact-checking retrieval module is also used to: when retrieving relevant knowledge base data, filter out the knowledge base data whose time is completely irrelevant to the corresponding time mark; when retrieving the core contradiction from the Internet to obtain a retrieval result, filter out the retrieval result whose time is completely irrelevant to the corresponding time mark.

[0008] Furthermore, the knowledge base retrieval module is also used to: crawl the data of the rumor refutation platform and establish a knowledge base.

[0009] Furthermore, the knowledge base retrieval module is also used to: Vectorize each piece of knowledge base data in the knowledge base to form vectorized knowledge base data; vectorize the entity relationship core contradiction and the entity attribute core contradiction to form a vectorized entity relationship core contradiction and a vectorized entity attribute core contradiction; Calculate the similarity score between each vectorized entity relationship core contradiction and each vectorized knowledge base data: ; wherein, represents the vectorized entity relationship core contradiction, represents the th piece of vectorized knowledge base data; Calculate the similarity score between each vectorized entity attribute core contradiction and each vectorized knowledge base data: ; Among them, represents the core contradiction of the vectorized entity attributes, represents the th vectorized knowledge base data; Take the larger value of the similarity score between the core contradiction of the vectorized entity relationship and the vectorized knowledge base data and the similarity score between the core contradiction of the vectorized entity attributes and the vectorized knowledge base data as the final similarity score of the corresponding knowledge base data: ; Select multiple knowledge base data with the highest final similarity scores as the relevant knowledge base data.

[0010] Furthermore, the Internet retrieval module is also used for: requiring the large language model to classify the credible retrieval results into information with no inspection value, information with inspection value, and completely irrelevant information in the third prompt word; the information with no inspection value involves entities related to the core contradiction but has no direct relationship with the core contradiction, or the provided information is useless, untrustworthy, or irrelevant; the information with inspection value is related to the core contradiction and provides useful information or clues for further investigation; the completely irrelevant information completely does not involve the entities mentioned in the core contradiction, or is completely irrelevant to the core contradiction, or is meaningless or misleading information.

[0011] Furthermore, the Internet retrieval module is also used for: establishing a source database and classifying the sources into high - credibility sources and low - credibility sources; the high - credibility sources include media and commercial websites; the low - credibility sources include individuals, self - media, and forums.

[0012] Furthermore, the rumor judgment and decision - making module is also used for: in the fifth prompt word: requiring the large language model to analyze whether the information in the judgment basis summary is consistent with the information in the text to be detected, and judge whether to affirm or deny the core contradiction; requiring the large language model to judge the reliability of the corresponding judgment basis based on the source of each judgment basis summary; requiring the large language model to judge whether the text to be detected is a rumor based on the reliability of the judgment basis summary and its impact on the core contradiction; requiring the large language model to output the analysis and judgment process of each judgment basis summary on the core contradiction of the text to be detected, and point out at least one of the most important judgment basis summaries in the analysis and judgment process.

[0013] A large - language - model - based rumor detection method using the above - mentioned system includes: Input the first prompt word and the text to be detected into the large language model, and extract the core contradiction of the entity relationship in the text to be detected. The core contradiction of the entity relationship is the contradiction between the mutual relationship of multiple entities in the text to be detected output by the large language model and the fact; input the second prompt word and the text to be detected into the large language model, and extract the core contradiction of the entity attribute in the text to be detected. The core contradiction of the entity attribute is the contradiction between the attribute of the entity in the text to be detected output by the large language model and the fact; Retrieve multiple pieces of knowledge base data most relevant to the core contradiction in the knowledge base through the RAG algorithm to obtain relevant knowledge base data; retrieve the core contradiction from the Internet to obtain a retrieval result, and filter the retrieval result through the source database to obtain a credible retrieval result; input the third prompt word, the core contradiction, and the credible retrieval result into the large language model to obtain a relevant credible retrieval result; the core contradiction includes the core contradiction of the entity relationship and the core contradiction of the entity attribute; Input the fourth prompt word, the core contradiction, and the judgment basis into the large language model, summarize each judgment basis, and extract the content that mutually corroborates or contradicts the core contradiction therein to form a judgment basis summary; the judgment basis includes the relevant knowledge base data and the relevant credible retrieval result; input the fifth prompt word, the text to be detected, the core contradiction, and the judgment basis summary into the large language model, and analyze the influence of each judgment basis summary on the core contradiction to form a detection result.

[0014] A computer-readable storage medium, the computer-readable storage medium includes a stored computer program, wherein when the computer program runs, it controls the device where the computer-readable storage medium is located to execute the above method.

[0015] A computer device, the computer device includes a memory, a processor, and a program stored and executable on the memory. When the program is executed by the processor, the above steps are implemented.

[0016] Beneficial effects: The present invention discloses a rumor detection system based on a large language model, including a core contradiction extraction module, a fact-checking retrieval module, and a rumor judgment and decision-making module. The core contradiction extraction module is used to extract the core contradiction of entity relationships and the core contradiction of entity attributes from the text to be detected through the large language model. The fact-checking retrieval module includes a knowledge base retrieval module and an Internet retrieval module, which are used to retrieve relevant credible judgment bases from the knowledge base and the Internet based on the core contradiction of entity relationships and the core contradiction of entity attributes. The rumor judgment and decision-making module is used to analyze and judge the impact of the judgment basis on the authenticity of the text to be detected through the large language model based on the core contradiction and the judgment basis, and finally output the analysis and demonstration and the detection result based on the judgment basis. By extracting the core contradiction through the large language model, the present invention can reduce the probability of being interfered by irrelevant information, facilitate subsequent retrieval, improve the information utilization efficiency, and enhance the rumor detection accuracy. At the same time, by retrieving relevant judgment bases on the knowledge base and the Internet through the fact-checking retrieval module, the source of the judgment basis is enriched, and at the same time, the hallucination of the large language model is avoided from affecting the detection result, and the reliability of the judgment basis is further improved through the filtering of the information source database. Finally, the rumor judgment and decision-making module enables the large language model to analyze the impact of each judgment basis summary on the core contradiction, form a detection result, simulates the process of human judgment, makes the detection result based on the judgment basis, and has interpretability. Description of the Drawings

[0017] In order to more clearly illustrate the technical solutions in the embodiments of the present application or the prior art, the following will briefly introduce the drawings required for use in the description of the embodiments or the prior art. Obviously, the following drawings are only some embodiments of the present application. For those of ordinary skill in the art, without creative efforts, other drawings can also be obtained based on these drawings.

[0018] Figure 1 It is a schematic block diagram of the method flow of the present invention; Figure 2 It is a schematic diagram of the method step flow of the present invention. Detailed Embodiments

[0019] In order to make the above-mentioned objects, features, and advantages of the present application more obvious and understandable, the following will make a detailed description of the specific embodiments of the present application in conjunction with the drawings. Many specific details are set forth in the following description in order to fully understand the present application. However, the present application can be implemented in many other ways different from those described herein. Those skilled in the art can make similar improvements without departing from the connotation of the present application. Therefore, the present application is not limited by the specific embodiments disclosed below.

[0020] In the description of this application, the terms "first" and "second" are only used for descriptive purposes and should not be construed as indicating or implying relative importance or implicitly specifying the quantity of the indicated technical features. Thus, features defined with "first" and "second" may explicitly or implicitly include at least one of such features. In the description of this application, the meaning of "a plurality" is at least two, such as two, three, etc., unless otherwise specifically defined.

[0021] In this application, unless otherwise clearly stipulated and defined, terms such as "installed", "connected", "linked", "fixed", etc. shall be understood in a broad sense. For example, it may be a fixed connection, a detachable connection, or integrated; it may be a mechanical connection or an electrical connection; it may be directly connected or indirectly connected through an intermediate medium, and it may be the communication inside two components or the interaction relationship between two components, unless otherwise clearly defined. For those of ordinary skill in the art, the specific meanings of the above terms in this application can be understood according to specific circumstances.

[0022] Embodiment 1: Referring to Figure 1 As shown, this embodiment provides a rumor detection system based on a large language model, including: A core contradiction extraction module, configured to: input a first prompt word and a text to be detected into the large language model to extract the core contradiction of entity relationships in the text to be detected, where the core contradiction of entity relationships is the contradiction between the mutual relationships of multiple entities in the text to be detected output by the large language model and the facts; input a second prompt word and the text to be detected into the large language model to extract the core contradiction of entity attributes in the text to be detected, where the core contradiction of entity attributes is the contradiction between the attributes of the entities in the text to be detected output by the large language model and the facts; Specifically, for the core contradiction of entity relationships, that is, the contradiction between the mutual relationships of multiple entities in the text to be detected and the facts, this embodiment designs the prompt words of the large language model to analyze the mutual relationships between different entities in the text to be detected, and then determines whether there is a possibility of exaggeration, distortion, or fabrication in these interaction relationships, and analyzes whether there is a possibility of inconsistency with the facts. In this embodiment, the first prompt word is: {The following text may be a rumor. Please identify two or more entities involved and their mutual relationships. Analyze whether the relationships between these entities conform to reality, whether there is any suspicion of exaggeration, distortion, or fabrication. Pay special attention to whether there are parts that do not conform to common sense or known facts. Ensure that all entities related to the contradiction are involved in the analysis, and clearly indicate the time when the contradiction occurs. Refine and summarize these parts into a simple sentence that can describe the contradiction, without summarizing whether it is reasonable:}; Input the above first prompt word and the text to be detected into the large language model, and the short sentence output by the large language model is the core contradiction of entity relationships ; For the core contradiction of entity attributes, that is, the contradiction between the attributes of the entity in the text to be detected and the facts, in this embodiment, prompt words for the large language model are designed to analyze and identify the key core entities in the text to be detected, and analyze whether there are exaggerations, distortions, or fabrications, and find the key points and key attributes that may be rumors. In this embodiment, the second prompt word is: {The following text may be a rumor. Please analyze the key entities mentioned in the text and their attribute descriptions, and judge whether these attributes conform to the actual situation and whether there are suspicions of exaggeration, distortion, or fabrication. Extract and summarize from these suspicious parts to form short sentences, ensuring that the entity and the corresponding state of the entity are described, and clearly indicating the time when the contradiction occurred, without summarizing whether it is reasonable:}; Input the above first prompt word and the text to be detected into the large language model, and the short sentence output by the large language model is the core contradiction of entity attributes .

[0023] The fact-checking retrieval module includes a knowledge base retrieval module and an Internet retrieval module; the knowledge base retrieval module is used to: retrieve multiple pieces of knowledge base data most relevant to the core contradiction in the knowledge base through the RAG algorithm to obtain relevant knowledge base data , and select k pieces of knowledge base data in this embodiment; the Internet retrieval module is used to: retrieve the core contradiction from the Internet to obtain a retrieval result, filter the retrieval result through the source database to obtain a credible retrieval result; input the third prompt word, the core contradiction, and the credible retrieval result into the large language model to obtain relevant credible retrieval results; the core contradiction includes the core contradiction of entity relationships and the core contradiction of entity attributes; In this embodiment, the knowledge base retrieval module is further used to: crawl the data of the rumor-refuting platform, which is the China Internet Joint Rumor Refuting Platform (www.piyao.org.cn) in this embodiment, and establish a knowledge base; The knowledge base retrieval module is further used to: Vectorize each piece of knowledge base data in the knowledge base to form vectorized knowledge base data ; vectorize the core contradiction of entity relationships and the core contradiction of entity attributes to form vectorized core contradiction of entity relationships and vectorized core contradiction of entity attributes; Calculate the similarity score between each vectorized core contradiction of entity relationships and each vectorized knowledge base data: ; Among them, represents the vectorized core contradiction of entity relationships, represents the th vectorized knowledge base data; Calculate the similarity scores between each vectorized entity attribute core contradiction and each vectorized knowledge base data: ; Among them, represents the vectorized entity attribute core contradiction, represents the th vectorized knowledge base data; Take the larger value of the similarity score between the vectorized entity relationship core contradiction and the vectorized knowledge base data and the similarity score between the vectorized entity attribute core contradiction and the vectorized knowledge base data as the final similarity score of the corresponding knowledge base data: ; Select multiple knowledge base data with the highest final similarity scores as the relevant knowledge base data; in this embodiment, select the top k knowledge base data with the highest final similarity scores as the relevant knowledge base data: ; Specifically, the Internet retrieval module is used to: call the search engine API interface to retrieve the core key contradiction and from the Internet, and obtain web pages or documents related to the core contradiction, that is, the retrieval results; read the web pages or documents given by the search engine one by one, filter out those from low - credibility information sources according to the information source database, and retain other search results from high - credibility information sources, that is, the credible retrieval results; Input the third prompt word, the core contradiction, and the credible retrieval results into the large - language model, so as to classify each credible retrieval result into three types: no inspection value, having inspection value, and completely irrelevant; no inspection value means that the searched content may involve the entity mentioned in the key contradiction, but has little relationship with the key contradiction; having inspection value means that the information is indeed related to the key contradiction and provides effective information worthy of further inspection; completely irrelevant means that the retrieval result does not involve the entity mentioned in the key contradiction at all, or is completely irrelevant to the key contradiction; in this embodiment, the third prompt word is: {The following is about a key contradiction <c>For the credible retrieval results, judge the relevance of each piece of information and classify it into three categories: 1. "Of no inspection value": This piece of information may involve the entity mentioned in the key contradiction, but has no direct relation to the key contradiction, or the information provided is useless, untrustworthy or irrelevant.

[0024] 2. "Of inspection value": This piece of information is related to the key contradiction and provides useful information or clues for further investigation.

[0025] 3. "Completely irrelevant": This piece of information does not involve the entity mentioned in the key contradiction at all, or is completely irrelevant to the key contradiction, and is even meaningless or misleading information.

[0026] Please classify each piece of information according to the following search results: Credible retrieval result 1: <Content of the credible retrieval result> Credible retrieval result 2: <Content of the credible retrieval result> …… Core contradiction c: <Core contradiction> Return the classification (of no inspection value, of inspection value, completely irrelevant) for each search result.}; Among them, the credible retrieval results of inspection value output by the large language model are the relevant credible retrieval results; reserve the number of relevant credible retrieval results for the core contradiction of entity relationship and the core contradiction of entity attribute respectively, that is, there are ; Merge the relevant knowledge base data and the relevant credible retrieval results to form the basis for judgment, that is, there are .

[0027] The rumor judgment and decision-making module is used to: input the fourth prompt word, the core contradiction, and the basis for judgment into the large language model, summarize each basis for judgment, and extract the content that mutually corroborates or contradicts the core contradiction therein to form a summary of the basis for judgment , in this embodiment, the summary of the basis for judgment marks the specific information source; the basis for judgment includes the relevant knowledge base data and the relevant credible retrieval results; input the fifth prompt word, the text to be detected, the core contradiction, and the summary of the basis for judgment into the large language model, and analyze the influence of each summary of the basis for judgment on the core contradiction to form a detection result; In this embodiment, the fourth prompt word is: {The following are multiple data entries from the rumor refutation knowledge base and the Internet, as well as the key contradiction and . Please summarize each piece of data and extract the part that may mutually corroborate or contradict the key contradiction therein. The generated summary should include the following content: 1. If the data conflicts with the key contradiction and is related to any one of them, analyze whether the data provides evidence to support or refute the contradiction. Clearly mark whether the data corroborates or contradicts the contradiction.

[0028] 2. Ensure that the information extracted in the abstract is clear and concise, only containing parts directly related to the key contradiction, and delete irrelevant information.

[0029] 3. Process each data entry to generate a short abstract, retaining the most critical information.

[0030] 4. Generate a final abstract dataset, including all processed entries, arranged in order.

[0031] List of data entries: - Data entry 1: <Data content 1> - Data entry 2: <Data content 2> ...... Key contradiction: -( ): <Core contradiction 1> -( ): <Core contradiction 2> Please output the abstract of each data entry item by item, ensuring that each abstract clearly indicates the relationship between the data and the contradiction, and output these abstracts in order:} The output of the large language model is the basis for judgment summary; Specifically, the rumor judgment decision module is further configured to: in the fifth prompt word: require the large language model to analyze whether the information in the judgment basis summary is consistent with the information in the text to be detected, and judge whether to affirm or deny the core contradiction; require the large language model to judge the reliability of the corresponding judgment basis based on the information source of each judgment basis summary; require the large language model to judge whether the text to be detected is a rumor based on the reliability of the judgment basis summary and the impact on the core contradiction; require the large language model to output the analysis and judgment process of each judgment basis summary on the core contradiction of the text to be detected, and point out at least one of the most important judgment basis summaries in the analysis and judgment process; In this embodiment, the fifth prompt word is: {Now it is necessary to judge whether a piece of information is a rumor. The following is the judgment basis summary from the rumor refutation knowledge base and the Internet , and the text to be detected that has been cleaned , and the core contradiction extracted from the text to be detected and 。Please determine whether this information is a rumor based on these judgment basis summaries and the text to be detected. Please make the judgment according to the following steps: 1. Information integration: Comprehensively consider all relevant content in the judgment basis summary and analyze whether they are consistent with the key information in the text to be detected , especially whether they support or refute the core contradiction mentioned before and .

[0032] 2. Contradiction verification: Compare with the core contradiction and to determine whether these data generate new contradictions and whether there are situations of information exaggeration, distortion or fabrication. If there are inconsistent or contradictory parts between the judgment basis summary and the text to be detected, focus on evaluating the impact of these parts on the rumor judgment.

[0033] 3. Information credibility: Analyze its impact on the rumor judgment based on the information source of each judgment basis summary, as well as the objectivity, authenticity, writing style, etc. of the text. Give priority to information with high credibility.

[0034] 4. Final judgment: Based on the above analysis, determine whether this text is a rumor. Please give the final judgment according to the following criteria: - If most of the summaries support the main content of the text to be detected and there are no obvious contradictions, it is determined as "not a rumor".

[0035] - If there is sufficient evidence indicating exaggeration, fabrication, distortion of facts or information inconsistent with objective reality, it is determined as "rumor".

[0036] - If the judgment basis summary and the text to be detected are inconsistent, but the contradiction is not significant, it is determined as "may be a rumor, further verification is required".

[0037] 5. Output of the analysis process: - Please output the detailed analysis process, including the reference content of each judgment basis summary, and point out which judgment basis summaries play a decisive role in the final judgment.

[0038] - Ensure to explain each judgment basis summary referred to and how they support or refute the core contradiction in the text to be detected.}.

[0039] In this embodiment, it also includes an original text preprocessing module for: removing the HTML tags and comments in the original text ; ; wherein represents the text after preprocessing and cleaning, defined as the text to be detected; It means removing HTML tags, and removing comments. In this embodiment, these two methods are completed based on regular expressions to avoid the influence of HTML tags and comments in the subsequent detection process.

[0040] As a further improvement of this embodiment, the core contradiction extraction module is further configured to: add a time mark when extracting the core contradiction of the entity relationship in the text to be detected, and add a time mark when extracting the core contradiction of the entity attribute in the text to be detected; the fact-checking retrieval module is further configured to: when retrieving relevant knowledge base data, filter out the knowledge base data whose time is completely irrelevant to the corresponding time mark; when retrieving the core contradiction from the Internet to obtain a retrieval result, filter out the retrieval result whose time is completely irrelevant to the corresponding time mark. In this embodiment, by adding a time mark to the core contradiction and filtering out irrelevant knowledge base data and Internet information, the influence of outdated information on the judgment and detection of rumors is avoided.

[0041] As a further improvement of this embodiment, the Internet retrieval module is further configured to: require the large language model to classify the credible retrieval results into information with no inspection value, information with inspection value, and completely irrelevant information in the third prompt word; the information with no inspection value involves the entity related to the core contradiction but has no direct relationship with the core contradiction, or the provided information is useless, untrustworthy, or irrelevant; the information with inspection value is related to the core contradiction and provides useful information or clues for further investigation; the completely irrelevant information completely does not involve the entity mentioned in the core contradiction, or is completely irrelevant to the core contradiction, or is meaningless or misleading information; In this embodiment, the Internet retrieval module is further configured to: establish a source database and divide the sources into high-credibility sources and low-credibility sources; among them, high-credibility sources have a certain credibility but may have some biases and mistakes. Usually, they are well-known media or institutions in the industry and have a certain degree of public credibility, including well-known mainstream media at home and abroad, reputable commercial websites, and industry media, etc.; low-credibility sources lack public credibility and may have biases, misleading, or exaggeration, including ordinary individuals on social media, forums of unknown origin, self-media accounts, etc. As a preference of this embodiment, the high-credibility sources include media and commercial websites; the low-credibility sources include individuals, self-media, and forums.

[0042] In the present embodiment, the large language models into which the first prompt word, the second prompt word, the third prompt word, the fourth prompt word, and the fifth prompt word are input can be the same or different large language models. As a preference of the present embodiment, the first prompt word, the second prompt word, the third prompt word, the fourth prompt word, and the fifth prompt word are input into different large language models, and these large language models are selected based on their proficient sub - fields or performances such as reasoning or summarization, so as to provide stronger performance for rumor detection. In the present embodiment, the large language model can be one or more of ChatGPT, BERT, Claude, LLaMA, Grok, DeepSeek, Gemini, Qwen, or can also be other large language models.

[0043] The present embodiment provides a rumor detection system based on a large language model, including a core contradiction extraction module, a fact - checking retrieval module, and a rumor judgment and decision - making module. The core contradiction extraction module is used to extract the core contradiction of entity relationships and the core contradiction of entity attributes from the text to be detected through the large language model. The fact - checking retrieval module includes a knowledge base retrieval module and an Internet retrieval module, and is used to retrieve relevant credible judgment bases from the knowledge base and the Internet based on the core contradiction of entity relationships and the core contradiction of entity attributes. The rumor judgment and decision - making module is used to analyze and judge the influence of the judgment bases on the authenticity of the text to be detected through the large language model based on the core contradiction and the judgment bases, and finally output an analysis and argumentation and a detection result based on the judgment bases. By extracting the core contradiction through the large language model in the present invention, the probability of being interfered by irrelevant information can be reduced, which is convenient for subsequent retrieval, improves the information utilization efficiency, and enhances the rumor detection accuracy; at the same time, by retrieving relevant judgment bases in the knowledge base and the Internet through the fact - checking retrieval module, the source of judgment bases is enriched, and the influence of hallucinations of the large language model on the detection result is avoided, and the reliability of the judgment bases is further improved through the filtering of the information source database; finally, the rumor judgment and decision - making module enables the large language model to analyze the influence of each judgment basis summary on the core contradiction to form a detection result, simulating the process of human judgment, making the detection result based on the judgment basis and having interpretability.

[0044] Embodiment Two: Refer to Figure 2 As shown, the present embodiment provides a large - language - model - based rumor detection method applying the system of Embodiment One, including: Input the first prompt word and the text to be detected into the large language model to extract the core contradiction of entity relationships in the text to be detected, where the core contradiction of entity relationships is the contradiction between the mutual relationships of multiple entities in the text to be detected output by the large language model and the facts; input the second prompt word and the text to be detected into the large language model to extract the core contradiction of entity attributes in the text to be detected, where the core contradiction of entity attributes is the contradiction between the attributes of the entities in the text to be detected output by the large language model and the facts; Retrieve multiple pieces of knowledge base data most relevant to the core contradiction in the knowledge base through the RAG algorithm to obtain relevant knowledge base data; retrieve the core contradiction from the Internet to obtain retrieval results, and filter the retrieval results through the information source database to obtain credible retrieval results; input the third prompt word, the core contradiction, and the credible retrieval results into the large language model to obtain relevant credible retrieval results; the core contradiction includes the core contradiction of entity relationships and the core contradiction of entity attributes. Input the fourth prompt word, the core contradiction, and the judgment basis into the large language model, summarize each judgment basis, and extract the content that mutually corroborates or contradicts the core contradiction therein to form a judgment basis summary; the judgment basis includes the relevant knowledge base data and the relevant credible retrieval results; input the fifth prompt word, the text to be detected, the core contradiction, and the judgment basis summary into the large language model, and analyze the impact of each judgment basis summary on the core contradiction to form a detection result.

[0045] Embodiment 3: This embodiment provides a computer-readable storage medium, which includes a stored computer program. When the computer program runs, it controls the device where the computer-readable storage medium is located to execute the method described in Embodiment 2.

[0046] Embodiment 4: This embodiment provides a computer device, which includes a memory, a processor, and a program stored and executable on the memory. When the program is executed by the processor, it implements the steps described in Embodiment 3.

[0047] The technical features of the above-described embodiments can be combined arbitrarily. For the sake of brevity of description, not all possible combinations of the technical features in the above embodiments are described. However, as long as there is no contradiction in the combination of these technical features, it should be considered as the scope described in this specification.

[0048] The above-described embodiments only represent several implementation manners of the present application. Their descriptions are relatively specific and detailed, but they should not be construed as limiting the scope of the patent application. It should be noted that for those of ordinary skill in the art, without departing from the concept of the present application, several modifications and improvements can still be made, and these all belong to the protection scope of the present application. Therefore, the protection scope of the patent of the present application should be subject to the appended claims.< / c>

Claims

1. A rumor detection system based on a large language model, characterized in that: include: The core contradiction extraction module is used to: input the first prompt word and the text to be detected into the large language model, and extract the core contradiction of the entity relationship in the text to be detected, wherein the core contradiction of the entity relationship is the contradiction between the mutual relationship of multiple entities in the text to be detected output by the large language model and the facts; Inputting the second prompt word and the text to be detected into the large language model, extracting the core contradiction of entity attributes in the text to be detected, wherein the core contradiction of entity attributes is the contradiction between the attributes of the entity in the text to be detected output by the large language model and the facts; The fact-checking retrieval module includes a knowledge base retrieval module and an Internet retrieval module; the knowledge base retrieval module is used to retrieve multiple pieces of knowledge base data most relevant to the core contradiction in the knowledge base through the RAG algorithm to obtain relevant knowledge base data; the Internet retrieval module is used to retrieve the core contradiction from the Internet to obtain a retrieval result, and filter the retrieval result through the information source database to obtain a credible retrieval result; Inputting the third prompt word, the core contradiction and the credible search result into the large language model to obtain relevant credible search results; the core contradiction includes the entity relationship core contradiction and the entity attribute core contradiction; The rumor judgment decision module is used to: input the fourth prompt word, core contradiction, and judgment basis into the large language model, summarize each judgment basis, and extract the content that mutually confirms or contradicts the core contradiction to form a judgment basis summary; the judgment basis includes the relevant knowledge base data and the relevant credible search results; input the fifth prompt word, the text to be detected, the core contradiction and the judgment basis summary into the large language model, analyze the impact of each judgment basis summary on the core contradiction, and form a detection result.

2. According to claim 1, a rumor detection system based on a large language model is characterized in that: The core contradiction extraction module is further used to: add a time mark when extracting the core contradiction of the entity relationship in the text to be detected, and add a time mark when extracting the core contradiction of the entity attribute in the text to be detected; The fact-checking retrieval module is also used to: when retrieving relevant knowledge base data, filter out knowledge base data whose time is completely irrelevant to the corresponding time mark; when retrieving core contradictions from the Internet to obtain retrieval results, filter out retrieval results whose time is completely irrelevant to the corresponding time mark.

3. The rumor detection system based on a large language model according to claim 1, characterized in that: The knowledge base retrieval module is also used to crawl rumor-refuting platform data and establish a knowledge base.

4. The rumor detection system based on a large language model according to claim 1, characterized in that: The knowledge base retrieval module is also used for: Vectorize each knowledge base data in the knowledge base to form vectorized knowledge base data; vectorize the entity relationship core contradiction and the entity attribute core contradiction to form vectorized entity relationship core contradiction and vectorized entity attribute core contradiction; Calculate the similarity score between each vectorized entity relationship core contradiction and each vectorized knowledge base data: ; in, Represents the core contradiction of vectorized entity relationships, Indicates Quantized knowledge base data; Calculate the similarity score between each vectorized entity attribute core contradiction and each vectorized knowledge base data: ; in, Represents the core contradiction of vectorized entity attributes, Indicates Quantized knowledge base data; The larger value of the similarity score between the core contradiction of the orientation quantified entity relationship and the vectorized knowledge base data and the similarity score between the core contradiction of the vectorized entity attribute and the vectorized knowledge base data is taken as the final similarity score of the corresponding knowledge base data: ; The multiple pieces of knowledge base data with the highest final similarity scores are selected as relevant knowledge base data.

5. The rumor detection system based on a large language model according to claim 1, characterized in that: The Internet search module is also used to: require the large language model in the third prompt word to divide the credible search results into information with no inspection value, information with inspection value and completely irrelevant information; the information with no inspection value is entities related to the core contradiction, but has no direct relationship with the core contradiction, or the information provided is useless, unreliable or irrelevant; the information with inspection value is related to the core contradiction and provides useful information or clues for further investigation; the completely irrelevant information is completely unrelated to the entities mentioned in the core contradiction, or is completely irrelevant to the core contradiction, or is meaningless or misleading information.

6. The rumor detection system based on a large language model according to claim 1, characterized in that: The Internet search module is also used to: establish a source database, and divide the sources into high-credibility sources and low-credibility sources; the high-credibility sources include media and commercial websites; the low-credibility sources include individuals, self-media and forums.

7. The rumor detection system based on a large language model according to claim 1, characterized in that: The rumor judgment decision module is further used to: in the fifth prompt word: require the large language model to analyze whether the information in the judgment basis summary is consistent with the information in the text to be detected, and judge whether to confirm or deny the core contradiction; The large language model is required to judge the reliability of the corresponding judgment basis based on the source of each judgment basis summary; The large language model is required to judge whether the text to be detected is a rumor based on the reliability of the judgment basis summary and the impact on the core contradiction; the large language model is required to output the analysis and judgment process of the core contradiction of the text to be detected for each judgment basis summary, and point out at least one of the most important judgment basis summaries in the analysis and judgment process.

8. A rumor detection method based on a large language model applied to the system as claimed in any one of claims 1 to 7, characterized in that: include: Inputting the first prompt word and the text to be detected into the large language model, extracting the core contradiction of entity relationship in the text to be detected, wherein the core contradiction of entity relationship is the contradiction between the mutual relationship of multiple entities in the text to be detected output by the large language model and the facts; Inputting the second prompt word and the text to be detected into the large language model, extracting the core contradiction of entity attributes in the text to be detected, wherein the core contradiction of entity attributes is the contradiction between the attributes of the entity in the text to be detected output by the large language model and the facts; Retrieve multiple pieces of knowledge base data most relevant to the core contradiction in the knowledge base through the RAG algorithm to obtain relevant knowledge base data; Retrieving core contradictions from the Internet to obtain search results, and filtering the search results through a source database to obtain credible search results; Inputting the third prompt word, the core contradiction and the credible search result into the large language model to obtain relevant credible search results; The core contradiction includes the entity relationship core contradiction and the entity attribute core contradiction; The fourth prompt word, the core contradiction, and the judgment basis are input into the large language model, each judgment basis is summarized, and the content that mutually confirms or contradicts the core contradiction is extracted to form a judgment basis summary; the judgment basis includes the relevant knowledge base data and the relevant credible search results; the fifth prompt word, the text to be detected, the core contradiction and the judgment basis summary are input into the large language model, the influence of each judgment basis summary on the core contradiction is analyzed, and the detection result is formed.

9. A computer-readable storage medium, characterized in that: The computer-readable storage medium includes a stored computer program, wherein when the computer program is executed, the device where the computer-readable storage medium is located is controlled to execute the method according to claim 8.

10. A computer device, characterized in that: The computer device comprises a memory, a processor and a program stored and executable on the memory, and the program implements the steps of the method according to claim 8 when executed by the processor.

Citation Information

Patent Citations

  • Multi-source and multi-mode fused knowledge reasoning method, system and device and medium

    CN119005340A

  • Information extraction method and device based on large language model, equipment and storage medium

    CN119415669A

  • Government affair processing method of big language model based on knowledge graph and electronic equipment

    CN119782545A

  • Information sending method and apparatus based on rumor prediction model, and computer device

    WO2022001517A1

Cited By

  • Manuscript generation service auxiliary system based on AI intelligent large model

    CN120781807A

  • Private network content copyright monitoring and evidence obtaining system based on AI and large model

    CN121637460A

  • Direct prompt injection threat mitigation using prompt processing units

    US20250292018A1