Log information desensitization method and device based on context semantics, equipment and medium

CN122113150APending Publication Date: 2026-05-29EASTERN COMM
View PDF 0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN Β· China
Patent Type
Applications(China)
Current Assignee / Owner
EASTERN COMM
Filing Date
2025-12-22
Publication Date
2026-05-29

AI Technical Summary

Technical Problem

Existing log anonymization schemes rely on static matching rules, lack an understanding of the log context semantics, leading to misjudgment or omission, failing to effectively identify sensitive information, and fixed anonymization rules result in over-anonymization or under-anonymization.

Method used

By converting log source files into a standardized log data structure, combining a sensitive information feature library and contextual correlation analysis, a weighted linear combination is used to evaluate keyword, delimiter, and positional features, and the desensitization strategy is dynamically adjusted to identify and process sensitive information.

Benefits of technology

It improves the accuracy and flexibility of sensitive information identification, reduces the probability of false positives and false negatives, achieves a balance between security and availability, and adapts to the desensitization needs of log data from multiple industries and of various types.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN122113150A_ABST
    Figure CN122113150A_ABST
Patent Text Reader

Abstract

The application relates to the technical field of intelligent decision-making, and specifically discloses a log information desensitization method and device based on context semantics, equipment and a medium, the method comprising the following steps: acquiring a log source file and converting the log source file into a standardized log data structure; inputting the log source file into a pre-configured sensitive information feature library for screening and calculation, so as to obtain at least one field name and context information; comprehensively evaluating keyword features, separator features and position features in a weighted linear combination mode, so as to obtain a context correlation degree score; determining a sensitive field according to a comparison result, and searching for associated sensitive information; and performing desensitization processing on the sensitive information according to a preset desensitization strategy, so as to obtain a desensitized log file. The present scheme realizes deep mining and analysis of context semantics, can more accurately determine whether a field name is a sensitive field and associated sensitive information, effectively reduces misjudgment and missed judgment problems caused by semantic loss, and improves the reliability of sensitive information identification.
Need to check novelty before this filing date? Find Prior Art