User information processing method fusing multi-source data
By obtaining and analyzing customs declaration templates in customs declaration information processing, determining the focus of filling, and mapping relationships with uploaded contents, the efficiency and accuracy of relationship establishment and customs declaration information processing are solved, and automated filling and information verification are realized.
Patent Information
- Application Number
- CN202510168402.0
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-02-17
- Publication Date
- 2025-06-06
AI Technical Summary
In customs declaration documents, how to establish a more accurate and clear relationship between multi-source data and process customs declaration information efficiently and accurately is still a technical problem.
Provide a user information processing method that integrates multi-source data. By obtaining customs declaration templates, analyzing templates, determining the filling focus; obtaining uploaded contents, mapping the relationship between the filling focus and the customs declaration templates, and obtaining the mapping results; and filling in customs declaration information based on the mapping results.
Through centralized template management and automated mapping processes, we ensure the accuracy and efficiency of customs declaration information, reduce manual errors and repeated work, and improve the accuracy and work efficiency of information matching.
Smart Images

Figure CN120106774A_ABST
Abstract
Description
Technical Field
[0001] The present application relates to the technical field of import and export management, and in particular to a user information processing method integrating multi-source data. Background Art
[0002] With the continuous development of global trade, the integration of multi-source data plays a vital role in the processing of customs declaration information. Through intelligent technology, accurate relationship mapping can be performed between different data sources, customs declaration information can be filled in automatically, manual operations can be reduced, and work efficiency and accuracy can be improved.
[0003] When customs declaration documents contain a large amount of complex information, how to establish more accurate and clear relationships between multi-source data and process customs declaration information efficiently and accurately remains a technical challenge. Summary of the invention
[0004] The present application provides a user information processing method for fusing multi-source data to solve the above problems.
[0005] In a first aspect, the present application provides a user information processing method for fusing multi-source data, the method comprising:
[0006] Obtain a customs declaration template, analyze the customs declaration template, and determine the key points to fill in;
[0007] Obtain the uploaded content, and perform relationship mapping between the uploaded content and the customs declaration template according to the filling key points to obtain a mapping result;
[0008] Fill in the customs declaration information according to the mapping result.
[0009] Through this solution, centralized template management ensures that the customs declaration templates obtained are the latest and comply with the latest regulations. Clarifying the key information fields in the customs declaration template provides a basis for automatic filling and information verification, reducing the workload of manual identification and filling. Quickly collect customs declaration documents provided by users to provide a data source for information extraction and analysis. Through the automated mapping process, the content uploaded by users is associated with the corresponding fields of the customs declaration template to improve the accuracy and efficiency of information matching. Automatic filling of customs declaration information is achieved, reducing manual entry errors and ensuring the accuracy of customs declaration information.
[0010] Optionally, analyzing the customs declaration template to determine the key points to be filled in includes:
[0011] Analyze the customs declaration template and determine the items to be filled in;
[0012] Identify the filled-in items and determine the item semantic information;
[0013] Determine the focus of filling in according to the semantic information of the project.
[0014] Through this plan, the key parts of the customs declaration information are clarified, providing a clear direction for data mapping. It is also clear which information must be filled in, as well as the specific location and format requirements of this information, providing a clear goal for information mapping.
[0015] Optionally, mapping the uploaded content with the customs declaration template according to the filling focus includes:
[0016] Obtain a preset upload template, and determine the upload file information according to the preset upload template;
[0017] Analyze the uploaded content to determine the file name of each uploaded file;
[0018] Matching the uploaded file information with the file name to obtain a matching result;
[0019] Determining a relationship file according to the matching result;
[0020] According to the item semantic information of the filling focus, the relationship file is relationally mapped with the filling focus.
[0021] Through this solution, we ensure that the declaration templates that comply with the latest regulations and standards are used, providing a standardized basis for filling in information. Clarifying which information is critical and needs to be processed first will help improve the efficiency and accuracy of information filling. Collect the declaration document data provided by users to provide necessary input for information processing. Through the mapping process, automated mapping reduces errors caused by human factors and improves the accuracy of data matching; reduces repeated manual inspections and entry work, and reduces labor intensity. Automated processing speeds up information mapping, especially when processing large amounts of data, and the efficiency is significantly improved. Generate mapping results to provide clear guidance for the automatic filling of declaration information and ensure the consistency and completeness of information. Automatically fill in declaration information to ensure the accuracy of information and the uniformity of format, and reduce subsequent review and modification work.
[0022] Optionally, mapping the relationship file with the filling focus according to the item semantic information of the filling focus includes:
[0023] Determine the detailed content of the relationship file according to the file name;
[0024] Analyze the detailed content according to the semantic information of the key items to determine several file areas containing the key items;
[0025] For each file area, the file area is analyzed to determine the coverage of the filling focus, and based on the coverage, several file areas are sorted, and the best mapping area is determined according to the sorting result, and the relationship between the best mapping area and the filling focus is mapped.
[0026] Through this solution, we ensure that the customs declaration templates that comply with the latest regulations are used, providing a standardized basis for filling in information. Clarifying which information is critical and providing direction for data extraction and mapping will help improve the accuracy and efficiency of information filling. Collect the customs declaration document data provided by users to provide necessary input for information processing. By analyzing the semantic information of the project, establish the logical association between the uploaded content and the fields of the customs declaration template to provide a basis for data mapping. Extract mapping rules, which define how to map the uploaded content to the fields of the customs declaration template, and provide specific operating guidelines for relationship mapping. Extract key information related to the key points of filling in the customs declaration template from the uploaded content to provide a data basis for the mapping process. Match the data items in the uploaded content with the fields of the customs declaration template to form a mapping relationship, laying the foundation for the automatic filling of customs declaration information. According to the mapping relationship, fill the data of the uploaded content into the corresponding fields of the customs declaration template to realize the automatic filling of customs declaration information.
[0027] Optionally, filling in customs declaration information according to the mapping result includes:
[0028] Analyze the optimal mapping area to determine the filling content of the filling focus;
[0029] After the filling is completed, the content coverage of the filling points is checked, and according to the check results, it is determined whether there are any missing fillings in the filling points;
[0030] If there is a missing entry, a full text analysis is performed on the file area outside the optimal mapping area;
[0031] According to the full text analysis results, the best filling content of the missing area is determined, and the missing area is filled with the best filling content.
[0032] Through this solution, critical information areas in customs declaration documents are identified by analyzing the best mapping areas. Determining the filling focus helps ensure that critical information in customs declaration documents is prioritized and reduce delays caused by missing information. Through content coverage check, verify whether key information has been filled in, thereby ensuring the basic integrity of customs declaration documents. Filling missing check can timely detect missing information and avoid customs declaration documents being unqualified due to missing information. Full-text analysis can extract missing information from non-optimal mapping areas and improve the integrity of customs declaration information. Through the full-text analysis results, the best filling content is found for the missing area to ensure the accuracy of the customs declaration information. Fill the best filling content into the missing area to ensure the integrity and accuracy of the customs declaration document.
[0033] Optionally, after obtaining the uploaded content, the step further includes:
[0034] Analyze the uploaded content to obtain content analysis results, and determine a number of document keywords based on the content analysis results and the filling focus;
[0035] For each document keyword, determining the file location of the document keyword according to the content analysis result;
[0036] According to the location of the file, a hidden number of the document keyword is set, and according to the hidden number, a jump link of the document keyword is established.
[0037] Through this solution, the structure of the uploaded file and the information it contains are identified through content analysis. It helps to identify key information points such as document number, date, amount, etc. in the customs declaration process, so that they can be quickly located and processed in subsequent steps. The specific location of key information in the original file is determined, which facilitates the operator to directly access and verify this information. Unique identifiers are assigned to each document keyword to facilitate tracking and management of this information while keeping the user interface tidy. A convenient navigation mechanism is created so that users can quickly jump to the part of the file containing key information, improving the efficiency of information retrieval.
[0038] Optionally, after filling the missing area with the best filling content, the method further includes:
[0039] Specially mark the missing area and obtain the hidden number of the best filled content;
[0040] Acquire a check signal, and determine whether the optimal filling content has been checked according to the check signal;
[0041] If it is checked, then according to the hidden number of the best filling content, the jump link of the best filling content is run to display all the contents of the location where the file of the best filling content is located.
[0042] Through this solution, through special markings, customs brokers can quickly identify areas of missing information in customs documents, improve work efficiency, and reduce the risk of missing important information. Hidden numbers provide a unique identifier for each best filling content, making it easy to track and manage filling suggestions while keeping the user interface tidy. The use of check signals allows customs brokers to know which filling content has been checked, helping to ensure that all information is properly reviewed. Ensure that only checked filling content will be further processed to avoid unverified information from being mistakenly adopted. The triggering and display functions of jump links allow customs brokers to directly access and view documents related to the best filling content, saving time in finding and verifying information.
[0043] Optionally, the method further includes:
[0044] Analyzing the inspection signal to determine actual operation instructions;
[0045] Determine whether there is a need for modification according to the actual operation instruction;
[0046] If it exists, the actual operation instruction is analyzed to determine the modified content, and according to the semantic information of the modified content and the uploaded content, a number of associated words and the file location of each associated word are determined;
[0047] Displaying the plurality of associated words and obtaining click feedback of the associated words;
[0048] The file display content is determined according to the click feedback of the associated word and the hidden number of the document keyword.
[0049] Through this solution, by analyzing the inspection signal, the user's intention can be identified, so as to perform the correct approval, rejection or modification of the customs declaration information. Identify which information needs further processing or correction, so as to ensure the accuracy and completeness of the customs declaration documents. By analyzing the modified content, the keywords (associated words) related to the modified content are identified, and the positions of these keywords in the original file are located, so that users can quickly find and view relevant information. By displaying the associated words, users can intuitively see which information needs attention, thereby improving the user's work efficiency. By obtaining the click feedback of the associated words, analyze which information the user is interested in or needs further processing, so as to provide more personalized services. According to the associated words and hidden numbers clicked by the user, determine the file content that should be displayed, and improve the efficiency of users in finding information. By displaying the relevant file content, users can analyze and verify the customs declaration information more comprehensively, thereby reducing errors and omissions.
[0050] Optionally, the performing full text analysis on the file area outside the optimal mapping area includes:
[0051] According to the inspection result, the filled content is determined;
[0052] According to the filled content, a full text analysis is performed on the file area outside the optimal mapping area.
[0053] Through this solution, identifying the filled content helps to avoid duplicate processing, improve efficiency, and ensure the accuracy of information. Preparation ensures the smooth progress of the full text analysis process. Extracting the full text content enables access and analysis of the entire document to discover important information that is not in the optimal mapping area. By cleaning and standardizing the text content, the quality and accuracy of the analysis are improved, preparing for subsequent information extraction and mapping. Identifying and extracting key information increases the completeness of customs declaration information and helps meet the detailed requirements of customs. Mapping the extracted information to the customs declaration template, even if the information is not in the optimal mapping area, it can be identified and processed by the system to ensure the comprehensiveness of the information. Verifying the mapped information ensures the accuracy of customs declaration information and reduces the risk of errors and violations. By inferring missing information, the completeness of customs declaration information is improved and delays caused by incomplete information are reduced. Supplementing the missing information in the customs declaration template ensures the completeness of the customs declaration document and avoids approval problems caused by incomplete information.
[0054] In a second aspect, the present application provides a user information processing system that integrates multi-source data, the system comprising:
[0055] The template analysis module is used to obtain the customs declaration template, analyze the customs declaration template, and determine the filling points;
[0056] A relationship mapping module is used to obtain the uploaded content, and perform relationship mapping between the uploaded content and the customs declaration template according to the filling key points to obtain a mapping result;
[0057] The information filling module is used to fill in the customs declaration information according to the mapping result.
[0058] Optionally, when the template analysis module analyzes the declaration template and determines the filling focus, it is used to:
[0059] Analyze the customs declaration template and determine the items to be filled in;
[0060] Identify the filled-in items and determine the item semantic information;
[0061] Determine the focus of filling in according to the semantic information of the project.
[0062] Optionally, when the relationship mapping module performs relationship mapping between the uploaded content and the customs declaration template according to the filling focus, it is used to:
[0063] Obtain a preset upload template, and determine the upload file information according to the preset upload template;
[0064] Analyze the uploaded content to determine the file name of each uploaded file;
[0065] Matching the uploaded file information with the file name to obtain a matching result;
[0066] Determining a relationship file according to the matching result;
[0067] According to the item semantic information of the filling focus, the relationship file is relationally mapped with the filling focus.
[0068] Optionally, when the relationship mapping module performs relationship mapping between the relationship file and the filling focus according to the item semantic information of the filling focus, it is used to:
[0069] Determine the detailed content of the relationship file according to the file name;
[0070] Analyze the detailed content according to the semantic information of the key items to determine several file areas containing the key items;
[0071] For each file area, the file area is analyzed to determine the coverage of the filling focus, and based on the coverage, several file areas are sorted, and the best mapping area is determined according to the sorting result, and the relationship between the best mapping area and the filling focus is mapped.
[0072] Optionally, when filling in the customs declaration information according to the mapping result, the information filling module is used to:
[0073] Analyze the optimal mapping area to determine the filling content of the filling focus;
[0074] After the filling is completed, the content coverage of the filling points is checked, and according to the check results, it is determined whether there are any missing fillings in the filling points;
[0075] If there is a missing entry, a full text analysis is performed on the file area outside the optimal mapping area;
[0076] According to the full text analysis results, the best filling content of the missing area is determined, and the missing area is filled with the best filling content.
[0077] Optionally, the user information processing system further includes a number generation module, which is used to:
[0078] Analyze the uploaded content to obtain content analysis results, and determine a number of document keywords based on the content analysis results and the filling focus;
[0079] For each document keyword, determining the file location of the document keyword according to the content analysis result;
[0080] According to the location of the file, a hidden number of the document keyword is set, and according to the hidden number, a jump link of the document keyword is established.
[0081] Optionally, the user information processing system further includes a jump display module, which is used to:
[0082] Specially mark the missing area and obtain the hidden number of the best filled content;
[0083] Acquire a check signal, and determine whether the optimal filling content has been checked according to the check signal;
[0084] If it is checked, then according to the hidden number of the best filling content, the jump link of the best filling content is run to display all the contents of the location where the file of the best filling content is located.
[0085] Optionally, the user information processing system further includes a modification display module, which is used to:
[0086] Analyzing the inspection signal to determine actual operation instructions;
[0087] Determine whether there is a need for modification according to the actual operation instruction;
[0088] If it exists, the actual operation instruction is analyzed to determine the modified content, and according to the semantic information of the modified content and the uploaded content, a number of associated words and the file location of each associated word are determined;
[0089] Displaying the plurality of associated words and obtaining click feedback of the associated words;
[0090] The file display content is determined according to the click feedback of the associated word and the hidden number of the document keyword.
[0091] Optionally, when the information filling module performs full text analysis on the file area outside the optimal mapping area, it is used to:
[0092] According to the inspection result, the filled content is determined;
[0093] According to the filled content, a full text analysis is performed on the file area outside the optimal mapping area. BRIEF DESCRIPTION OF THE DRAWINGS
[0094] In order to more clearly illustrate the embodiments of the present application or the technical solutions in the prior art, a brief introduction will be given below to the drawings required for use in the embodiments or the description of the prior art. Obviously, the drawings described below are some embodiments of the present application. For ordinary technicians in this field, other drawings can be obtained based on these drawings without paying any creative labor.
[0095] Figure 1 A schematic diagram of an application scenario provided for an embodiment of the present application;
[0096] Figure 2 A flowchart of a user information processing method for fusing multi-source data provided in one embodiment of the present application;
[0097] Figure 3 A schematic diagram of the structure of a user information processing system for fusing multi-source data provided in one embodiment of the present application. DETAILED DESCRIPTION
[0098] In order to make the purpose, technical scheme and advantages of the embodiments of the present application clearer, the technical scheme in the embodiments of the present application will be clearly and completely described below in conjunction with the drawings in the embodiments of the present application. Obviously, the described embodiments are part of the embodiments of the present application, not all of the embodiments. Based on the embodiments in the present application, all other embodiments obtained by ordinary technicians in this field without creative work are within the scope of protection of this application.
[0099] In addition, the term "and / or" in this article is only a description of the association relationship of associated objects, indicating that there can be three relationships. For example, A and / or B can represent: A exists alone, A and B exist at the same time, and B exists alone. In addition, the character " / " in this article, unless otherwise specified, generally means that the associated objects before and after are in an "or" relationship.
[0100] The embodiments of the present application are further described in detail below in conjunction with the drawings in the specification.
[0101] With the continuous development of global trade, the integration of multi-source data plays a vital role in the processing of customs declaration information. Through intelligent technology, accurate relationship mapping can be performed between different data sources, customs declaration information can be filled in automatically, manual operations can be reduced, and work efficiency and accuracy can be improved.
[0102] When customs declaration documents contain a large amount of complex information, how to establish more accurate and clear relationships between multi-source data and process customs declaration information efficiently and accurately remains a technical challenge.
[0103] Based on this, the present application provides a user information processing method that integrates multi-source data, obtains a customs declaration template, analyzes the customs declaration template, and determines the filling focus; obtains the uploaded content, and maps the uploaded content with the customs declaration template according to the filling focus to obtain the mapping result; fills in the customs declaration information according to the mapping result. Centralized template management ensures that the customs declaration template obtained is the latest and complies with the latest regulations. Clarify the key information fields in the customs declaration template to provide a basis for automatic filling and information verification, and reduce the workload of manual identification and filling. Quickly collect the customs declaration documents provided by the user to provide a data source for information extraction and analysis. Through the automated mapping process, the content uploaded by the user is associated with the corresponding fields of the customs declaration template to improve the accuracy and efficiency of information matching. Automatic filling of customs declaration information is achieved, reducing errors in manual entry and ensuring the accuracy of customs declaration information.
[0104] Figure 1 A schematic diagram of an application scenario provided for the present application, in which the method provided by the present application is applied. Specifically, the method provided by the present application is applied to any server, and the server interacts with the customs declaration system and the user end, extracts the customs declaration template from the customs declaration system, and ensures that the customs declaration template obtained is the latest and complies with the latest regulations. Clarify the key information fields in the customs declaration template, provide a basis for automatic filling and information verification, and reduce the workload of manual identification and filling. Receive the uploaded content uploaded by the user on the user end, quickly collect the customs declaration documents provided by the user, and provide a data source for information extraction and analysis. Through the automated mapping process, the content uploaded by the user is associated with the corresponding fields of the customs declaration template to improve the accuracy and efficiency of information matching. Automatic filling of customs declaration information is achieved, errors in manual entry are reduced, and the accuracy of customs declaration information is ensured. For specific implementation methods, please refer to the following embodiments.
[0105] Figure 2 This is a flowchart of a user information processing method for integrating multi-source data provided by an embodiment of the present application. The method of this embodiment can be applied to the server in the above scenario. Figure 2 As shown, the method includes:
[0106] S201, obtain the customs declaration template, analyze the customs declaration template, and determine the key points to be filled in;
[0107] A customs declaration template can be a standardized electronic document or data structure.
[0108] Emphasis can be information fields or areas in the customs declaration documents that require special attention and accurate filling.
[0109] Specifically, the customs declaration template is extracted from the customs declaration system. The template text is parsed using natural language processing technology to extract the template structure information. The key fields and required items in the template are analyzed to identify the key points of filling. The semantic analysis technology is used to analyze the semantic information of each filled-in item to determine its importance.
[0110] S202, obtaining the uploaded content, and mapping the uploaded content with the customs declaration template according to the filling key points to obtain the mapping result;
[0111] The uploaded content can be relevant documents and information submitted by the user for filling in and processing customs declaration information.
[0112] Relationship mapping can be the process of matching and associating the content uploaded by the user with the corresponding fields in the customs declaration template during the customs declaration information processing.
[0113] The mapping result may be the output obtained after the relationship mapping is completed.
[0114] Specifically, receive the uploaded content uploaded by the user on the user side. Pre-process the uploaded file. Extract the key information in the uploaded file according to the preset upload template rules. Use text matching and pattern recognition technology to match the uploaded content with the filling key points of the customs declaration template. Determine the mapping relationship through the algorithm and generate the mapping result.
[0115] S203. Fill in the customs declaration information according to the mapping result.
[0116] Specifically, according to the mapping results, the relevant information in the customs declaration template is automatically filled in. The filled-in information is verified to ensure that it meets the customs declaration requirements.
[0117] Through this solution, centralized template management ensures that the customs declaration templates obtained are the latest and comply with the latest regulations. Clarifying the key information fields in the customs declaration template provides a basis for automatic filling and information verification, reducing the workload of manual identification and filling. Quickly collect customs declaration documents provided by users to provide a data source for information extraction and analysis. Through the automated mapping process, the content uploaded by users is associated with the corresponding fields of the customs declaration template to improve the accuracy and efficiency of information matching. Automatic filling of customs declaration information is achieved, reducing manual entry errors and ensuring the accuracy of customs declaration information.
[0118] In some embodiments, the customs declaration template is analyzed to determine the items to be filled in; the items to be filled in are identified to determine the semantic information of the items; and the filling focus is determined based on the semantic information of the items.
[0119] Fill-in items can be information items that need to be filled in the customs declaration template and are an important part of a complete customs declaration document.
[0120] Item semantic information may be specific information and rules related to each field in the customs declaration template, including the data type of the field, whether it is a required field, the value range, the logical relationship between fields, etc.
[0121] Specifically, the pre-defined customs declaration templates are extracted from the database. Based on the structural information of the templates, natural language processing technology is used to identify the key fields that need to be filled in. The semantics of these fields are analyzed to determine their filling requirements and priorities.
[0122] Through this plan, the key parts of the customs declaration information are clarified, providing a clear direction for data mapping. It is also clear which information must be filled in, as well as the specific location and format requirements of this information, providing a clear goal for information mapping.
[0123] In some embodiments, a preset upload template is obtained, and the uploaded file information is determined based on the preset upload template; the uploaded content is analyzed to determine the file name of each uploaded file; the uploaded file information is matched with the file name to obtain a matching result; based on the matching result, a relationship file is determined; and based on the project semantic information of the filled-in key points, a relationship mapping is performed between the relationship file and the filled-in key points.
[0124] A preset upload template can be a predefined file format or structure.
[0125] The uploaded file information may be the data and content contained in the file uploaded by the user.
[0126] The file name can be the identification name used to upload the file.
[0127] The matching result may be the result obtained by matching the information in the uploaded file with the fields in the customs declaration template during the customs declaration information processing.
[0128] A relationship file can be a file used to record and describe the relationship between different data elements during the customs declaration information processing process.
[0129] Specifically, retrieve or download the corresponding customs declaration template file from the customs declaration system. Parse the structure of the customs declaration template, extract field labels and filling instructions. Mark the required fields and key information fields in the template as filling points. Receive files uploaded by users. Convert the uploaded files to text format for subsequent processing. Clean the text to remove format information and irrelevant content. Use text analysis technology to extract key data points in the uploaded file. According to the filling points of the customs declaration template, identify the corresponding information in the uploaded content. Formulate mapping rules based on the fields of the customs declaration template and the key information of the uploaded content. Formulate mapping rules such as field matching logic and data conversion rules. Apply mapping rules to match the extracted uploaded content with the fields of the customs declaration template. For the fields that are successfully matched, record the mapping relationship. Check the mapping results to confirm whether all filling points have been correctly mapped. Mark the anomalies or omissions found during the mapping process. According to the verification results, adjust the mapping rules to improve the mapping accuracy. Prepare a mapping report, the process and results of the successfully mapped fields and the parts that require manual intervention.
[0130] Through this solution, we ensure that the declaration templates that comply with the latest regulations and standards are used, providing a standardized basis for filling in information. Clarifying which information is critical and needs to be processed first will help improve the efficiency and accuracy of information filling. Collect the declaration document data provided by users to provide necessary input for information processing. Through the mapping process, automated mapping reduces errors caused by human factors and improves the accuracy of data matching; reduces repeated manual inspections and entry work, and reduces labor intensity. Automated processing speeds up information mapping, especially when processing large amounts of data, and the efficiency is significantly improved. Generate mapping results to provide clear guidance for the automatic filling of declaration information and ensure the consistency and completeness of information. Automatically fill in declaration information to ensure the accuracy of information and the uniformity of format, and reduce subsequent review and modification work.
[0131] In some embodiments, the detailed content of the relationship file is determined based on the file name; the detailed content is analyzed based on the project semantic information of the filling focus to determine several file areas containing the filling focus; for each file area, the file area is analyzed to determine the coverage of the filling focus, and based on the coverage, several file areas are sorted, and the best mapping area is determined based on the sorting result, and the best mapping area is relationally mapped with the filling focus.
[0132] The details may be the specific information contained in the uploaded file or document during the customs declaration process.
[0133] The file region can be divided into different parts or regions when processing the uploaded file, and each region contains information matching the mapping relationship type.
[0134] An ordering structure can be a way of organizing data or information by arranging the data according to specific rules for easy searching, analysis, and processing.
[0135] A coverage scope can be a relationship file or the set of data fields to which a mapping rule applies.
[0136] The best mapping area may be a file area that can maximize mapping accuracy and efficiency, determined according to mapping rules and algorithms during the relationship mapping process.
[0137] The sorting result can be the result obtained after sorting a group of numbers, strings, objects and other data.
[0138] Specifically, read the customs declaration template file and extract all fields and their attributes in the template. According to the parsing results of the customs declaration template, determine each field that needs to be filled in and its project semantic information. Read the relationship file containing mapping rules. Parse the mapping rules in the relationship file to extract the mapping relationship between fields, data conversion rules and verification rules. Extract key information related to filling in the key points through text analysis, data mining and other technologies. According to the project semantic information and the rules in the relationship file, use string matching, pattern recognition and other technologies to match the data items in the uploaded content with the fields in the customs declaration template. According to the mapping rules, perform necessary format conversion or unit conversion on the successfully matched data. Verify the mapped data to ensure that the data meets the customs declaration requirements, fill in the missing data or mark it as requiring manual intervention. Apply the established mapping relationship to fill the data in the uploaded content into the corresponding fields of the customs declaration template.
[0139] Through this solution, we ensure that the customs declaration templates that comply with the latest regulations are used, providing a standardized basis for filling in information. Clarifying which information is critical and providing direction for data extraction and mapping will help improve the accuracy and efficiency of information filling. Collect the customs declaration document data provided by users to provide necessary input for information processing. By analyzing the semantic information of the project, establish the logical association between the uploaded content and the fields of the customs declaration template to provide a basis for data mapping. Extract mapping rules, which define how to map the uploaded content to the fields of the customs declaration template, and provide specific operating guidelines for relationship mapping. Extract key information related to the key points of filling in the customs declaration template from the uploaded content to provide a data basis for the mapping process. Match the data items in the uploaded content with the fields of the customs declaration template to form a mapping relationship, laying the foundation for the automatic filling of customs declaration information. According to the mapping relationship, fill the data of the uploaded content into the corresponding fields of the customs declaration template to realize the automatic filling of customs declaration information.
[0140] In some embodiments, the best mapping area is analyzed to determine the content to be filled in the key points; after the filling is completed, a content coverage check is performed on the key points to determine whether there are any missing fillings in the key points based on the check results; if there are any missing fillings, a full-text analysis is performed on the file area outside the best mapping area; based on the full-text analysis results, the best filling content for the missing area is determined, and the best filling content is filled in the missing area.
[0141] The content to be filled in can be the specific information items filled in the customs declaration template based on the uploaded file information.
[0142] Content coverage check can be a check on the completeness and consistency of the filled content during the process of filling in the customs declaration information.
[0143] The inspection result can be the result of a review of the accuracy, completeness and compliance of the filled-in information after the customs declaration information is completed.
[0144] Missing fields may be fields that are identified as unfilled or incompletely filled in during the customs declaration information filling process.
[0145] The full-text analysis result may be a result obtained after full-text analysis of unmapped areas in the customs declaration document.
[0146] The best filling content may be the best filling data automatically identified and recommended according to the mapping rules and customs declaration requirements.
[0147] Missing areas may be fields or areas that are detected to be missing necessary information during the customs declaration information filling process.
[0148] Specifically, determine which areas in the customs declaration template are the best mapping areas. These areas are usually standardized information filling areas. Analyze the best mapping areas and determine which information is the focus of filling. These key information are usually the key information that must be filled in the customs declaration documents. Perform content coverage check on the filling points to ensure that all key information has been filled in. Based on the results of the content coverage check, determine whether there are any missing fillings. If there are missing fillings, perform a full text analysis of the document areas outside the best mapping areas to find content containing missing information. Based on the results of the full text analysis, determine the best filling content for the missing areas. Fill in the missing areas with the determined best filling content to ensure the integrity of the customs declaration documents.
[0149] Through this solution, critical information areas in customs declaration documents are identified by analyzing the best mapping areas. Determining the filling focus helps ensure that critical information in customs declaration documents is prioritized and reduce delays caused by missing information. Through content coverage check, verify whether key information has been filled in, thereby ensuring the basic integrity of customs declaration documents. Filling missing check can timely detect missing information and avoid customs declaration documents being unqualified due to missing information. Full-text analysis can extract missing information from non-optimal mapping areas and improve the integrity of customs declaration information. Through the full-text analysis results, the best filling content is found for the missing area to ensure the accuracy of the customs declaration information. Fill the best filling content into the missing area to ensure the integrity and accuracy of the customs declaration document.
[0150] In some embodiments, the uploaded content is analyzed to obtain content analysis results, and a number of document keywords are determined based on the content analysis results and filling priorities; for each document keyword, the file location of the document keyword is determined based on the content analysis results; based on the file location, a hidden number of the document keyword is set, and based on the hidden number, a jump link for the document keyword is established.
[0151] The content analysis result may be the output obtained after analyzing the content of the uploaded customs declaration document.
[0152] Document keywords can be key words that appear frequently in customs declaration documents and play a guiding role in information extraction.
[0153] The file location can be a physical or virtual path where the file is stored.
[0154] A hidden number may be a number that exists in the customs declaration document but is not directly displayed to the user.
[0155] A jump link can be a hyperlink in a customs declaration document or related system that connects to other related documents or information.
[0156] Specifically, perform text analysis on the uploaded content and use NLP technology to extract key information. Identify and classify different information types in the documents, such as commercial invoices, packing lists, and transport documents. Determine the keywords for each document based on the content analysis results and the key points of filling. Determine the location of this information in the original uploaded file based on the document keywords. If the uploaded content contains multiple files, identify the key document information contained in each file. Assign a unique hidden number to each document keyword for identification and tracking. The hidden number can be generated based on the text hash or serial number attribute of the keyword. Use the hidden number to create a jump link for each document keyword so that you can quickly locate the corresponding file location when processing customs declaration information. Implement the jump function so that users can directly access the relevant document information by clicking on the link.
[0157] Through this solution, the structure of the uploaded file and the information it contains are identified through content analysis. It helps to identify key information points such as document number, date, amount, etc. in the customs declaration process, so that they can be quickly located and processed in subsequent steps. The specific location of key information in the original file is determined, which facilitates the operator to directly access and verify this information. Unique identifiers are assigned to each document keyword to facilitate tracking and management of this information while keeping the user interface tidy. A convenient navigation mechanism is created so that users can quickly jump to the part of the file containing key information, improving the efficiency of information retrieval.
[0158] In some embodiments, the missing area is specially marked, and the hidden number of the best filled content is obtained; a check signal is obtained, and based on the check signal, it is determined whether the best filled content is checked; if checked, based on the hidden number of the best filled content, the jump link of the best filled content is run to display all the contents of the file location of the best filled content.
[0159] The check signal can be a signal used to indicate whether a certain step or condition is met during the customs declaration information processing.
[0160] Specifically, the uploaded file is parsed to extract information from text and images. The extracted content is analyzed to identify missing areas in the customs declaration document. The missing areas are specially marked, highlighted with different colors or with specific symbols. Based on contextual information and historical data, the best filling content is recommended for the missing area. A hidden number is assigned to the recommended filling content for tracking and reference. An inspection signal is obtained from the user interface or automated system, which indicates that the user will check the recommended filling content. Based on the inspection signal, it is determined whether the recommended filling content has been checked. If the content has been checked, the inspection result is recorded. If the recommended filling content is checked, the corresponding jump link is triggered according to the hidden number. All the content at the location of the file pointed to by the jump link is displayed for user viewing and verification.
[0161] Through this solution, through special markings, customs brokers can quickly identify areas of missing information in customs documents, improve work efficiency, and reduce the risk of missing important information. Hidden numbers provide a unique identifier for each best filling content, making it easy to track and manage filling suggestions while keeping the user interface tidy. The use of check signals allows customs brokers to know which filling content has been checked, helping to ensure that all information is properly reviewed. Ensure that only checked filling content will be further processed to avoid unverified information from being mistakenly adopted. The triggering and display functions of jump links allow customs brokers to directly access and view documents related to the best filling content, saving time in finding and verifying information.
[0162] In some embodiments, the inspection signal is analyzed to determine the actual operation instruction; based on the actual operation instruction, it is determined whether there is a modification requirement; if so, the actual operation instruction is analyzed to determine the modification content, and based on the semantic information of the modification content and the uploaded content, a number of associated words and the file location of each associated word are determined; a number of associated words are displayed, and click feedback on the associated words is obtained; based on the click feedback on the associated words and the hidden number of the document keyword, the file display content is determined.
[0163] The actual operation instruction may be a specific operation command given by the user to instruct the customs declaration information processing system to perform a specific task.
[0164] A modification request may be a request to change the information found to be inaccurate, incomplete or non-compliant during the preparation or review of customs declaration documents.
[0165] The modified content may be specific file information or data that needs to be changed.
[0166] Semantic information can be meaningful content contained in the customs declaration document. Associated words can be words used to connect different concepts or information in the customs declaration document or related text.
[0167] The document location can be a specific place where the customs declaration documents are stored or placed.
[0168] The associated word click feedback may be a response or feedback given after the user clicks the associated word during interaction.
[0169] The document presentation content may be the information visually presented in the document.
[0170] Specifically, parse the inspection signal provided by the user to determine the actual operation instruction type such as "approved", "need to be modified", "view details", etc. Based on the actual operation instructions, determine whether there is a need to modify the customs declaration information. If there is a need for modification, further analyze the actual operation instructions to determine the specific modification content. According to the semantic information of the modified content, combined with the uploaded content, identify the keywords (associated words) related to the modified content. Determine the position of each associated word in the original uploaded file, involving multiple files. Display the identified associated words to the user, usually highlighted on the user interface or displayed in a list form. Track the user's click behavior on the associated words, and obtain feedback such as which associated word is clicked and how many times it is clicked. Determine the file content that should be displayed based on the associated word click feedback and the hidden number of the document keyword. If the user clicks on a certain associated word, the file content related to the associated word will be displayed.
[0171] Through this solution, by analyzing the inspection signal, the user's intention can be identified, so as to perform the correct approval, rejection or modification of the customs declaration information. Identify which information needs further processing or correction, so as to ensure the accuracy and completeness of the customs declaration documents. By analyzing the modified content, the keywords (associated words) related to the modified content are identified, and the positions of these keywords in the original file are located, so that users can quickly find and view relevant information. By displaying the associated words, users can intuitively see which information needs attention, thereby improving the user's work efficiency. By obtaining the click feedback of the associated words, analyze which information the user is interested in or needs further processing, so as to provide more personalized services. According to the associated words and hidden numbers clicked by the user, determine the file content that should be displayed, and improve the efficiency of users in finding information. By displaying the relevant file content, users can analyze and verify the customs declaration information more comprehensively, thereby reducing errors and omissions.
[0172] In some embodiments, performing a full text analysis on the file area outside the optimal mapping area includes: determining the filled content according to the inspection result; and performing a full text analysis on the file area outside the optimal mapping area according to the filled content.
[0173] Filled content can be information that has been entered by the user or automatically filled in the customs declaration document or electronic form.
[0174] Full-text analysis can be the process of parsing and analyzing unstructured text in customs declaration documents using natural language processing technology to extract key information.
[0175] Specifically, analyze the inspection results and identify the filled content areas in the customs declaration documents. Prepare text processing tools and algorithms for full-text analysis. Extract all text content from the documents that is not in the optimal mapping area. Clean and standardize the extracted text content to remove noise, segment words, remove stop words, etc. Use NLP technology to identify and extract key information such as product name, quantity, price, etc. in the text. Map the extracted information to the corresponding fields in the customs declaration template. Verify the mapped information to ensure its accuracy and compliance. If some information is missing, try to infer the missing information through context or related data. Based on the inference results, supplement the missing information in the customs declaration template.
[0176] Through this solution, identifying the filled content helps to avoid duplicate processing, improve efficiency, and ensure the accuracy of information. Preparation ensures the smooth progress of the full text analysis process. Extracting the full text content enables access and analysis of the entire document to discover important information that is not in the optimal mapping area. By cleaning and standardizing the text content, the quality and accuracy of the analysis are improved, preparing for subsequent information extraction and mapping. Identifying and extracting key information increases the completeness of customs declaration information and helps meet the detailed requirements of customs. Mapping the extracted information to the customs declaration template, even if the information is not in the optimal mapping area, it can be identified and processed by the system to ensure the comprehensiveness of the information. Verifying the mapped information ensures the accuracy of customs declaration information and reduces the risk of errors and violations. By inferring missing information, the completeness of customs declaration information is improved and delays caused by incomplete information are reduced. Supplementing the missing information in the customs declaration template ensures the completeness of the customs declaration document and avoids approval problems caused by incomplete information.
[0177] Figure 3 A schematic diagram of the structure of a user information processing system that integrates multi-source data provided by an embodiment of the present application is shown in FIG. Figure 3 As shown, the user information processing system 300 for fusing multi-source data of this embodiment includes:
[0178] The template analysis module 301 is used to obtain the customs declaration template, analyze the customs declaration template, and determine the filling points;
[0179] The relationship mapping module 302 is used to obtain the uploaded content, and perform relationship mapping between the uploaded content and the customs declaration template according to the filling key points to obtain a mapping result;
[0180] The information filling module 303 is used to fill in the customs declaration information according to the mapping result.
[0181] Optionally, the template analysis module 301 analyzes the declaration template to determine the filling focus, and is used to:
[0182] Analyze the customs declaration template and determine the items to be filled in;
[0183] Identify the filled-in items and determine the item semantic information;
[0184] Determine the focus of filling in according to the semantic information of the project.
[0185] Optionally, when the relationship mapping module 302 performs relationship mapping between the uploaded content and the customs declaration template according to the filling focus, it is used to:
[0186] Obtain a preset upload template, and determine the upload file information according to the preset upload template;
[0187] Analyze the uploaded content to determine the file name of each uploaded file;
[0188] Matching the uploaded file information with the file name to obtain a matching result;
[0189] Determining a relationship file according to the matching result;
[0190] According to the item semantic information of the filling focus, the relationship file is relationally mapped with the filling focus.
[0191] Optionally, when the relationship mapping module 302 performs relationship mapping between the relationship file and the filling focus according to the project semantic information of the filling focus, it is used to:
[0192] Determine the detailed content of the relationship file according to the file name;
[0193] Analyze the detailed content according to the semantic information of the key items to determine several file areas containing the key items;
[0194] For each file area, the file area is analyzed to determine the coverage of the filling focus, and based on the coverage, several file areas are sorted, and the best mapping area is determined according to the sorting result, and the relationship between the best mapping area and the filling focus is mapped.
[0195] Optionally, when filling in the customs declaration information according to the mapping result, the information filling module 303 is used to:
[0196] Analyze the optimal mapping area to determine the filling content of the filling focus;
[0197] After the filling is completed, the content coverage of the filling points is checked, and according to the check results, it is determined whether there are any missing fillings in the filling points;
[0198] If there is a missing entry, a full text analysis is performed on the file area outside the optimal mapping area;
[0199] According to the full text analysis results, the best filling content of the missing area is determined, and the missing area is filled with the best filling content.
[0200] Optionally, the user information processing system 300 further includes a number generation module 304, which is used to:
[0201] Analyze the uploaded content to obtain content analysis results, and determine a number of document keywords based on the content analysis results and the filling focus;
[0202] For each document keyword, determining the file location of the document keyword according to the content analysis result;
[0203] According to the location of the file, a hidden number of the document keyword is set, and according to the hidden number, a jump link of the document keyword is established.
[0204] Optionally, the user information processing system 300 further includes a jump display module 305, which is used to:
[0205] Specially mark the missing area and obtain the hidden number of the best filled content;
[0206] Acquire a check signal, and determine whether the optimal filling content has been checked according to the check signal;
[0207] If it is checked, then according to the hidden number of the best filling content, the jump link of the best filling content is run to display all the contents of the location where the file of the best filling content is located.
[0208] Optionally, the user information processing system 300 further includes a modification display module 306, which is used to:
[0209] Analyzing the inspection signal to determine actual operation instructions;
[0210] Determine whether there is a need for modification according to the actual operation instruction;
[0211] If it exists, the actual operation instruction is analyzed to determine the modified content, and according to the semantic information of the modified content and the uploaded content, a number of associated words and the file location of each associated word are determined;
[0212] Displaying the plurality of associated words and obtaining click feedback of the associated words;
[0213] The file display content is determined according to the click feedback of the associated word and the hidden number of the document keyword.
[0214] Optionally, when the information filling module 303 performs full text analysis on the file area outside the optimal mapping area, it is used to:
[0215] According to the inspection result, the filled content is determined;
[0216] According to the filled content, a full text analysis is performed on the file area outside the optimal mapping area.
[0217] The system of this embodiment can be used to execute the method of any of the above embodiments. The implementation principles and technical effects are similar and will not be described in detail here.
Claims
1. A user information processing method for integrating multi-source data, characterized in that: include: Obtain a customs declaration template, analyze the customs declaration template, and determine the key points to fill in; Obtain the uploaded content, and perform relationship mapping between the uploaded content and the customs declaration template according to the filling key points to obtain a mapping result; Fill in the customs declaration information according to the mapping result.
2. The method according to claim 1, characterized in that The analysis of the customs declaration template and determination of the key points to be filled in include: Analyze the customs declaration template and determine the items to be filled in; Identify the filled-in items and determine the item semantic information; Determine the focus of filling in according to the semantic information of the project.
3. The method according to claim 2, characterized in that According to the filling focus, mapping the uploaded content with the customs declaration template includes: Obtain a preset upload template, and determine the upload file information according to the preset upload template; Analyze the uploaded content to determine the file name of each uploaded file; Matching the uploaded file information with the file name to obtain a matching result; Determining a relationship file according to the matching result; According to the item semantic information of the filling focus, the relationship file is relationally mapped with the filling focus.
4. The method according to claim 3, characterized in that The mapping of the relationship file with the filling focus according to the item semantic information of the filling focus includes: Determine the detailed content of the relationship file according to the file name; Analyze the detailed content according to the semantic information of the key items to determine several file areas containing the key items; For each file area, the file area is analyzed to determine the coverage of the filling focus, and based on the coverage, several file areas are sorted, and the best mapping area is determined according to the sorting result, and the relationship between the best mapping area and the filling focus is mapped.
5. The method according to claim 4, characterized in that Filling in the customs declaration information according to the mapping result includes: Analyze the optimal mapping area to determine the filling content of the filling focus; After the filling is completed, the content coverage of the filling points is checked, and according to the check results, it is determined whether there are any missing fillings in the filling points; If there is a missing entry, a full text analysis is performed on the file area outside the optimal mapping area; According to the full text analysis results, the best filling content of the missing area is determined, and the missing area is filled with the best filling content.
6. The method according to claim 5, characterized in that After obtaining the uploaded content, the method further includes: Analyze the uploaded content to obtain content analysis results, and determine a number of document keywords based on the content analysis results and the filling focus; For each document keyword, determining the file location of the document keyword according to the content analysis result; According to the location of the file, a hidden number of the document keyword is set, and according to the hidden number, a jump link of the document keyword is established.
7. The method according to claim 6, characterized in that After filling the missing area with the best filling content, the method further includes: Specially mark the missing area and obtain the hidden number of the best filled content; Acquire a check signal, and determine whether the optimal filling content has been checked according to the check signal; If it is checked, then according to the hidden number of the best filling content, the jump link of the best filling content is run to display all the contents of the location where the file of the best filling content is located.
8. The method according to claim 7, characterized in that The method further comprises: Analyzing the inspection signal to determine actual operation instructions; Determine whether there is a need for modification according to the actual operation instruction; If it exists, the actual operation instruction is analyzed to determine the modified content, and according to the semantic information of the modified content and the uploaded content, a number of associated words and the file location of each associated word are determined; Displaying the plurality of associated words and obtaining click feedback of the associated words; The file display content is determined according to the click feedback of the associated word and the hidden number of the document keyword.
9. The method according to claim 5, characterized in that The full text analysis of the file area outside the optimal mapping area includes: According to the inspection result, the filled content is determined; According to the filled content, a full text analysis is performed on the file area outside the optimal mapping area.
10. A user information processing system integrating multi-source data, characterized in that: include: The template analysis module is used to obtain the customs declaration template, analyze the customs declaration template, and determine the key points to be filled in; A relationship mapping module is used to obtain the uploaded content, and perform relationship mapping between the uploaded content and the customs declaration template according to the filling key points to obtain a mapping result; The information filling module is used to fill in the customs declaration information according to the mapping result.
Citation Information
Patent Citations
Accounting information processing method and device based on voice recognition and electronic equipment
CN110659970A
Customs declaration information management method
CN114581055A
Original document management method and system
CN116050366A
Online customs declaration generation method and system
CN118095229A