File acquisition method, device, computer device and storage medium
By receiving file call requests, obtaining file encoding and project identification, determining file scheduling time, and generating file management information, it solves the problem of difficult to manage complex multi-project file acquisition requirements in traditional technology, and realizes efficient and refined file management and acquisition.
Patent Information
- Application Number
- CN202111393958.8
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2021-11-23
- Publication Date
- 2025-06-27
- Estimated Expiration
- 2041-11-23
AI Technical Summary
Traditional technology is difficult to effectively manage complex and multi-project file acquisition requirements, resulting in high workload for designers and managers, and the inability to achieve batch planning and refined management of requirements.
Provide a file acquisition method, which obtains file encoding and project identification by receiving file call requests, determines the file scheduling time, obtains the target template based on this, fills the template to generate file management information, which is used to indicate the correspondence between the file scheduling time, the project identification and the file to be extracted, and obtains the corresponding file at a specified time.
It realizes refined control of the file acquisition requirements of multiple projects, reduces file acquisition time, improves efficiency, reduces the number of duplicate file acquisition times, and realizes efficient and refined file management.
Smart Images

Figure CN114036187B_ABST
Abstract
Description
Technical Field
[0001] This application relates to the technical field of file acquisition, and particularly to a file acquisition method, device, computer device, and storage medium. Background Art
[0002] With the increase in the number of projects and specialties, there are personalized file acquisition requirements for different projects within the same specialty. Among these requirements, some are standardized data such as fixed patterns, fixed templates, and fixed parameters, while others are personalized data. To avoid the infinite expansion of personalized data and reduce the workload of designers and requirements managers, it is necessary to perform standardized classification and processing, merge some common, representative, and similar requirements into one standardized requirement, achieve unified control, reduce the number of personalized requirements, and at the same time ensure that the standardized requirements have different planned time requirements for different projects.
[0003] In traditional technologies, the focus is mainly on the macro level. Different data is stored in different databases, and different identities are used to operate the data in different databases to avoid data misoperations and chaos during information submission.
[0004] However, in current traditional methods, when the number of files to be acquired reaches a new level, a more effective way of control is needed at this time. It is impossible to solve the unified management of complex and multi-project requirements, and the workload of designers and managers remains high. It is impossible to achieve batch planning and refined management of requirements. Summary of the Invention
[0005] Based on this, in view of the above technical problems, it is necessary to provide a file acquisition method, device, computer device, and storage medium that can perform refined control.
[0006] A file acquisition method, the method includes:
[0007] Receiving a file call request, obtaining the file code and project identifier of the file to be extracted carried by the file call request, obtaining the category to which the file code belongs, and obtaining the file scheduling time mapped to the category to which the file code belongs;
[0008] Based on the file call request, obtaining a target template set, selecting a template identifier from the template set according to the file call request, and combining the template fields corresponding to the selected template identifier to form a target template;
[0009] Filling the target template with the file scheduling time, the project identifier, and the file code to obtain file management information, where the file management information is used to indicate the correspondence between the file scheduling time, the project identifier, and the file to be extracted;
[0010] At the file scheduling time, following the corresponding relationship, obtain the corresponding file to be extracted according to the project identifier.
[0011] A file acquisition device, the device includes:
[0012] An acquisition time determination module, configured to receive a file call request, obtain the file code and project identifier of the file to be extracted carried in the file call request, obtain the category to which the file code belongs, and obtain the file scheduling time mapped to the category to which the file code belongs;
[0013] A template acquisition module, configured to obtain a target template set based on the file call request, select a template identifier in the template set according to the file call request, and combine the template fields corresponding to the selected template identifier to form a target template;
[0014] A corresponding relationship determination module, configured to fill the target template with the file scheduling time, the project identifier, and the file code to obtain file management information, where the file management information is used to indicate the corresponding relationship between the file scheduling time, the project, and the file to be extracted;
[0015] A file extraction module, configured to, at the file scheduling time, following the corresponding relationship, obtain the corresponding file to be extracted according to the project.
[0016] A computer device, including a memory and a processor, where the memory stores a computer program, and when the processor executes the computer program, the following steps are implemented:
[0017] Receive a file call request, obtain the file code and project identifier of the file to be extracted carried in the file call request, obtain the category to which the file code belongs, and obtain the file scheduling time mapped to the category to which the file code belongs;
[0018] Obtain a target template set based on the file call request, select a template identifier in the template set according to the file call request, and combine the template fields corresponding to the selected template identifier to form a target template;
[0019] Fill the target template with the file scheduling time, the project identifier, and the file code to obtain file management information, where the file management information is used to indicate the corresponding relationship between the file scheduling time, the project identifier, and the file to be extracted;
[0020] At the file scheduling time, following the corresponding relationship, obtain the corresponding file to be extracted according to the project identifier.
[0021] A computer-readable storage medium, on which a computer program is stored, and when the computer program is executed by a processor, the following steps are implemented:
[0022] Receive a file call request, obtain the file code and project identifier of the file to be extracted carried by the file call request, obtain the category to which the file code belongs, and obtain the file scheduling time mapped to the category to which the file code belongs;
[0023] Based on the file call request, obtain a target template set, select a template identifier from the template set according to the file call request, and combine the template fields corresponding to the selected template identifier to form a target template;
[0024] Fill the target template with the file scheduling time, the project identifier, and the file code to obtain file management information, which is used to indicate the correspondence between the file scheduling time, the project identifier, and the file to be extracted;
[0025] At the file scheduling time, follow the correspondence to obtain the corresponding file to be extracted according to the project identifier.
[0026] The above file acquisition method, device, computer device, and storage medium, the category of the file to be extracted, determine the scheduling time of the file to be extracted, reorganize the file code data required by each project, aggregate the scattered data into specific categories, reduce the time for obtaining files, and improve the relevant efficiency; and by selecting the template identifier in the template set and combining the fields of each template to obtain a target template, its compatibility can be increased, the reuse of templates can be realized, the compatibility of the present application can be enhanced, the repeated process of developing templates can be reduced, and the relevant efficiency can be improved. Then, through the filled target template, the correspondence and scheduling time of the files required by multiple projects can be obtained at the same time, and at the file scheduling time corresponding to the category of the file to be extracted, the file to be extracted corresponding to the project can be obtained through the correspondence, so as to obtain scattered files more efficiently, reduce the number of times of obtaining duplicate files, and achieve efficient and refined management. Description of the Drawings
[0027] Figure 1 It is an application environment diagram of the file acquisition method in an embodiment;
[0028] Figure 2 It is a flowchart of the file acquisition method in an embodiment;
[0029] Figure 3 It is a flowchart of obtaining the file scheduling time in an embodiment;
[0030] Figure 4 It is a flowchart of forming a target template in another embodiment;
[0031] Figure 5 It is a flowchart of obtaining a target template in an embodiment;
[0032] Figure 6 It is a schematic flow diagram of obtaining file management information in an embodiment;
[0033] Figure 7 It is a schematic flow diagram of correcting file management information in an embodiment;
[0034] Figure 8 It is a schematic flow diagram of forming intermediate data in an embodiment;
[0035] Figure 9 It is a schematic flow diagram of forming a first file acquisition template in an embodiment;
[0036] Figure 10 It is a structural block diagram of a file acquisition device in an embodiment;
[0037] Figure 11 It is an internal structure diagram of a computer device in an embodiment. Specific embodiments
[0038] In order to make the objectives, technical solutions and advantages of the present application clearer, the present application will be further described in detail below with reference to the accompanying drawings and embodiments. It should be understood that the specific embodiments described herein are only used to explain the present application and are not used to limit the present application.
[0039] The file acquisition method provided by the present application can be applied to an application environment as shown in Figure 1 After the terminal 102 sends a file call request, the server 104 receives the file call request, obtains the file code and project identifier of the file to be extracted carried by the file call request, obtains the category to which the file code belongs, and obtains the file scheduling time mapped to the category of the file code; based on the file call request, obtain a target template set, select a template identifier from the template set according to the file call request, combine the template fields corresponding to the selected template identifier to form a target template; fill the target template with the file scheduling time, the project identifier and the file code to obtain file management information, which is used to indicate the correspondence between the file scheduling time, the project identifier and the file to be extracted; at the file scheduling time, follow the correspondence and obtain the corresponding file to be extracted according to the project identifier.
[0040] Among them, the terminal 102 can be, but is not limited to, various personal computers, laptop computers, smart phones, tablet computers and portable wearable devices, and the server 104 can be implemented by an independent server or a server cluster composed of multiple servers.
[0041] In one embodiment, Figure 2 As shown, a file acquisition method is provided, which is applied to Figure 1 The server in the example is used to illustrate the following steps:
[0042] Step 202, receiving a file call request, obtaining the file code and project identifier of the file to be extracted carried in the file call request, obtaining the category to which the file code belongs, and obtaining the file scheduling time mapped to the category to which the file code belongs.
[0043] The file call request can be a request instruction directly generated by some means, or it can be generated by some functions and corresponding conditional expressions. Among them, the file call request can be issued by the account that manages the files; it can also be directly generated by some mapping relationship after the files required by one or some projects meet the preset file acquisition conditions. The preset conditions can be dynamically adjusted, and the preset conditions can be any factors such as the number and priority of the required files of one or more projects.
[0044] File coding, also known as job coding, is used to identify the file to be extracted and can be used to classify files. It can be a general identifier for a specific file to reduce the number of mappings and avoid data redundancy. It can also be a file extraction identifier corresponding to a general identifier, and data is obtained through the file extraction identifier to facilitate the process of obtaining files. The file coding can be the description information of the file to be extracted, or it can correspond to or belong to the description information of the file to be extracted. The description information of the file to be extracted can also be a combination of multiple description information of the file to be extracted, such as extraction major, extraction department, item number, extraction classification, unit number, job code, template code, etc., to further increase the efficiency of file acquisition.
[0045] The category to which the file code belongs is used to identify the scheduling time of the file to be extracted. It can be set according to the file priority, the project priority, the file similarity, the correlation between files, the range of the file code, or the completed publishing batch of the file. The completed publishing batch of the file is used to characterize the publishing time of a certain type of file. The publishing time can be a specific time, a range of time values, or a mapped value.
[0046] The file scheduling time can be used to remind professionals of the deadline for obtaining project files. Professionals can be one or more of the following: the person who provides files, the person who requires files, the person who coordinates and plans, etc. The file scheduling time can be the deadline for preparing files, the deadline for the funding process node, and / or the deadline for solidifying the process.
[0047] In an alternative embodiment, the method includes the step of generating a file call request, which includes the step of importing a standardized list: The account for information submission management can add a standardized list item by item manually, or add one or more related data such as a standardized list through batch import in EXCEL; the standardized list mainly includes the project description information of the file to be extracted and the file description information to be extracted, and the standardized list can correspond to standardized general fields such as the design stage of the project, sub-item, system, information submission specialty, information submission department, item number, information collection specialty, information collection department, key information submission, material name, requirement description, information submission classification, unit number, operation code, template code, etc., and the standardized general fields are applicable to at least some projects. Among them, the standardized list is similar to a template in the conventional sense, and the standardized general fields are at least some fields of the standardized list.
[0048] In an alternative embodiment, the account for file acquisition management issues a file call request, obtains multiple project codes associated with the file call request, and then obtains the category of the file to be extracted to which the project code belongs, so as to obtain the file scheduling time mapped by the category of the file to be extracted.
[0049] Step 204, obtain a target template set based on the file call request, select a template identifier in the template set according to the file call request, and combine the template fields corresponding to the selected template identifier to form a target template.
[0050] The template set includes at least one template group, template or one template field; at least one template, at least one template field or at least one template identifier associated with the template group can be included in one template group; at least one template field or at least one template identifier associated with the template can be included in one template; the template field can be a general field such as the design stage of the project, project identifier, file code of the file to be extracted, etc., or a special field specifically corresponding to one or more fields or projects, and the special field can be used to characterize the associated template field associated with the field.
[0051] The target template is composed of at least one template. The template constituting the target template can be the above-mentioned standardized list, or the template associated with the above-mentioned standardized list can be included therein. The role of the target template is similar to a cornerstone and is the basis for constructing the corresponding relationship between the project and the file to be obtained.
[0052] In the step of obtaining the target template set based on the file call request, other types of corresponding relationships can be realized according to mapping relationships such as dependence, association, combination, etc., or relationships obtained by calculation of certain functions or models. Among them, one file call request can correspond to one or more template sets, and one template set can also correspond to one or more file call requests.
[0053] In the step of selecting the template identifier in the template set according to the file call request and combining the template fields corresponding to the selected template identifier to form the target template, one or more template identifiers can be selected through the template code, template identifier or certain mapping relationships in the file call request, and the combined target template can be obtained. Alternatively, the above-mentioned standardized list can be directly used as the template without selecting the template identifier.
[0054] Step 206: Obtain the project identifier, and use the file scheduling time, project identifier, and file code to fill the target template to obtain the file management information, which is used to indicate the correspondence between the file scheduling time, project, and files to be extracted.
[0055] The project identifier is used to represent the project that requires files to be extracted. The project identifier itself can be project description information, or it can belong to or correspond to project description information. The project description information can be a combination of information such as the design stage, sub-item, system, information collection specialty, information collection department, and requirement description of the project, so as to clarify the categories of data extraction and improve the efficiency of information acquisition.
[0056] The file management information is the information used to manage the file acquisition process. It can contain at least part of the project description information and file description information to establish the correspondence between the project and the files to be extracted. It can also include the file scheduling time, which is used to indicate the time for obtaining files between the project and the files to be extracted. The file management information can be the filled target template or certain information mapped by the filled target template. When using the above-mentioned standardized list as the main template, the generated file management information can be called the project planning list, and the process of generating the file management information can be called list planning.
[0057] In an optional embodiment, in the step of using the file scheduling time, project identifier, and file code to fill the target template, these information can be directly filled in, or the information corresponding to the project identifier and file code or the information set to which they belong can be filled into the target template. For example: in the process of using the project description information and the description information of the files to be extracted to fill the target template, according to the fields of the target template, the project description information and the description information of the files to be extracted can be input into the corresponding fields to directly establish the correspondence between the description information; or through the fields of the target template, the mapping relationship between the project description information and the description information of the files to be extracted can be input into the corresponding fields to indirectly establish the correspondence between the description information; or according to the fields of the target template, the project identifier and the file code of the files to be extracted can be input into the corresponding fields, and then the corresponding description information can be obtained according to the data corresponding to the project identifier and the file code of the files to be extracted.
[0058] Step 208, at the file scheduling time, follow the corresponding relationship and obtain the corresponding files to be extracted according to the project.
[0059] To solve problems such as overly scattered data, duplicate data in different projects, and a large workload in managing information submission requirements, different categories of files are extracted to achieve efficient file extraction and reduce the number of file extractions. Thus, discrete files are aggregated and output at a specific time, which can reduce the number of times of obtaining duplicate files.
[0060] In the above file acquisition method, according to the categories of files to be extracted, the scheduling time of the files to be extracted is determined, so that the file encoding data required by each project is reorganized, and the scattered data is aggregated into specific categories, reducing the time for obtaining files and improving relevant efficiency; by selecting the template identifier in the template set and combining the fields of each template to obtain the target template, its compatibility can be increased, the reuse of the template can be realized, the compatibility of this application can be enhanced, the repeated process of developing templates can be reduced, and relevant efficiency can be improved. Then, through the filled target template, the corresponding relationship and scheduling time of the files required by multiple projects are obtained at the same time, and at the file scheduling time corresponding to the category of files to be extracted, the files to be extracted corresponding to the project are obtained through the corresponding relationship, and the scattered files can be obtained more efficiently, reducing the number of times of obtaining duplicate files, and realizing efficient and refined management.
[0061] In one embodiment, as Figure 3 shown, obtaining the category to which the file code belongs and obtaining the file scheduling time mapped to by the category to which the file code belongs includes:
[0062] Step 302, obtain the job classification mapping table, and according to the job classification mapping table, determine the category of the file to be extracted to which the file code belongs.
[0063] The job classification mapping table can be a mapping table or any data structure that can achieve the same function as the mapping table. It is used to represent the corresponding relationship between key-value pairs. Among them, the description information of the file to be extracted is the key, and the category to which the description information of the file to be extracted belongs is the value for constructing the description information of the file to be extracted.
[0064] According to the job classification mapping table, determining the category of the file to be extracted to which the file code belongs can be implemented in various ways: it can be mapped through Map or mapped using other algorithms. For example, it can be implemented by using algorithms such as hashmap algorithm, TreeMap algorithm, set method, or by using a queue in combination with a pointer.
[0065] Step 304, obtain the period mapping table corresponding to the category of the file to be extracted, and based on the period mapping table, estimate the activity period corresponding to the category of the file to be extracted.
[0066] The cycle mapping table contains the time information corresponding to each document category. Different document categories correspond to different document scheduling times, which can be time periods of different lengths or different time points. The data in the cycle mapping table can be obtained from historical data or estimated through models. The cycle mapping table can be a mapping table or any data structure that can achieve the same function as a mapping table, and it is used to represent the correspondence between key-value pairs. In the cycle mapping table, the document category is the key, and the active cycle is the value corresponding to the document category.
[0067] In an optional embodiment, the active cycle includes one or more of the first cycle, the second cycle, and the third cycle; wherein, the first cycle is the time to complete the first file to be extracted, which is an imperfect file and generally not within the fields of the above-mentioned standardized list, i.e., the FRE time; the second cycle is the time to complete the first file to be extracted, which belongs to a relatively complete version and status, i.e., the FIN time; the third cycle is the time to complete the third file to be extracted, which belongs to the solidified version, i.e., the FRZ time. Among them, the FIN time and the FRZ time are different fields in the standardized list.
[0068] Step 306: Obtain the initial time corresponding to the file code, and the initial time is the time when the file call request is received.
[0069] The initial time is the time when the file call request is received, and the time when the file call request is received can be a time point or a time period. For example: the initial time can be the computer time when the file call request is received, or it can be within the range of the computer time.
[0070] Step 308: Calculate based on the initial time and the estimated active cycle to obtain the file scheduling time.
[0071] Calculating based on the initial time and the active cycle can be a calculation between two time periods or a calculation between two time points. For example: the active cycle is the time period before or within 3 months, and the initial time can be a larger time period such as January, a smaller time period such as January 1st, or a time point such as 1:01:01 on January 1st; based on the same argument, when the active cycle is set to the time point of 1:01:01 on March 1st, the initial time can also be a larger time period such as January, a smaller time period such as January 1st, or a time point such as 1:01:01 on January 1st.
[0072] In this embodiment, first, the category of the file to be extracted to which the file code belongs is determined through the job classification mapping table, and then the acquisition time corresponding to different file categories is determined through the cycle mapping table. Thus, through the two-step mapping method, not only the amount of data for a single mapping is reduced, the data is more finely controlled, but also the problem of the same data being acquired multiple times is avoided, improving the file acquisition efficiency. Even for large-scale projects to acquire a large number of files, refined management can be achieved.
[0073] In one embodiment, as Figure 4 shown, select the template identifier in the template set according to the file call request, and combine the template fields corresponding to the selected template identifier to form a target template, including:
[0074] Step 402, obtain the template code in the file call request, and match the template code with the template identifiers in the template set.
[0075] The template code is the data carried by the file call request and is used to represent the template required by the file call request. Different templates include at least partially different fields. The template code can be of any data type. It can be a numerical type such as integer, long integer or floating point type, or it can be a string type or a boolean type. The template code can be any part of the template. It can be the template identifier or another template field.
[0076] Matching the template code with the template identifier means collecting templates in the template set and determining whether there is a template identifier corresponding to the template code. There are multiple implementation methods for this step. It can directly compare the template code with the template identifier to determine whether they are the same, so as to reduce the difficulty of building the data system; it can compare the template code with the fields of each template to determine whether there is a correlation, and then determine the field corresponding to the template code. According to the field corresponding to the template code, the template identifier of the template to which the corresponding field belongs is determined as the template identifier corresponding to the template code.
[0077] Step 404, if the match is successful, then according to the matched template identifier, obtain the template fields corresponding to the matched template identifier from the template set, and combine the selected template fields to obtain the target template.
[0078] According to user requirements, the method for constructing the target template can be selected. For example: according to the matched template identifier, obtain the template to be spliced corresponding to the matched template identifier, and splice each template to be spliced to generate the target template to ensure the integrity of the template and avoid data omission or loss; or the fields of the template to be spliced can be extracted, and redundant fields can be selectively removed, and then the target template can be generated to reduce the total amount of data and the calculation amount; or a certain template can be used as the default template, and the templates related to the default template can be selected for combination.
[0079] Optionally, after matching the template encoding with the template identifier, it further includes: if the matching fails, obtain a default template. The default template may be the above-mentioned standardized list.
[0080] In this embodiment, by matching the template encoding with the template identifier, selecting one or more templates for combination to generate the required target template, highly flexible template combination can be achieved, with high compatibility, which can meet the requirements of multiple fields, multiple systems, and multiple projects, and realize the reuse of template data.
[0081] In one embodiment, as Figure 5 shown, the template identifier includes an associated template identifier. According to the matched template identifier, obtain the template fields corresponding to the matched template identifier from the template set, and combine the selected template fields. The obtained target template includes:
[0082] Step 502, obtain a first file acquisition template, where the first file acquisition template includes a first description information field and an associated template identifier field.
[0083] The first file acquisition template is a default template, which may be the above-mentioned standardized list or a template with other fields. The first file acquisition template includes a first description information field, and the first description information field is a common field. The common field may be a common field of at least two existing templates or a field solidified by other means.
[0084] In an optional embodiment, obtaining the first file acquisition template includes: in response to receiving a file call request, obtain the first file acquisition template. In this embodiment, regardless of whether the template encoding matches the template identifier, the first file acquisition template can be obtained, realizing standardization at the basic level, reducing the number of matches, and reducing the calculation amount.
[0085] In an optional embodiment, obtaining the first file acquisition template includes: after receiving the file call request, detect whether there is a template encoding corresponding to the associated template identifier; if detected, obtain the first file acquisition template. Among them, after the template encoding matches the template identifier, the first file acquisition template can be obtained, otherwise other templates can be used, enriching the types of templates, with stronger compatibility and higher standardization degree.
[0086] Step 504, fill the matched associated template identifier into the associated template identifier field to obtain a target associated identifier.
[0087] In the process of obtaining the target associated identifier, the corresponding relationship between the first file acquisition template and the second file acquisition template is constructed. And the associated template identifier field may include any number of associated template identifiers to realize the combination of multiple templates.
[0088] Step 506, obtain a second file acquisition template corresponding to the target association identifier, where the second file acquisition template includes a second description information field.
[0089] The second file acquisition template is a template associated with the first file acquisition template. The second description information field in it may not have a difference in the degree of generality from the first description field to achieve composite scheduling of the templates; the second description information field may also be any template, which can be set arbitrarily to meet diverse requirements.
[0090] Step 508, add the second description information field to the first file acquisition template to generate a target template.
[0091] The target template includes a first description information field, a second description information field, and an associated template identifier field. Among them, the first description information field is a general field, which is a field common to multiple fields, systems, majors, or projects; the second description information field is a field focusing on personalization and has high flexibility; the associated template identifier field is used to select a second file acquisition template to be associated with the first file acquisition template to achieve combination and reuse of the templates.
[0092] In the step of adding the second description information field to the first file acquisition template, any combination of the first description information field and the second description information field can be used, and the combined description information field is used as the field of the target template; or the order of obtaining information for the first description information field and the second description information field can be set to achieve information acquisition with different priorities.
[0093] Optionally, if the target association identifier includes multiple associated template identifiers, multiple second description information fields can be added to the first file acquisition template, and redundant fields of multiple first file acquisition templates can be eliminated to reduce the total amount of data obtained and avoid repeated data acquisition.
[0094] In this embodiment, by setting the first description information field as the field of a preset target template and obtaining a pending second description information field through the associated template identifier field, a target template is constructed, taking into account both coarse-grained and fine-grained data control and accurately controlling the file acquisition process.
[0095] In an optional embodiment, the method further includes a step of generating a first description information field, and the steps include:
[0096] Obtain historical description information fields, and obtain duplicate description information fields from the historical description information fields;
[0097] Review the duplicate description information fields, and use the duplicate project information fields that pass the review as the first description information fields.
[0098] A historical description information field, which can be set according to the description information in the database or directly generated according to certain conditions, indicators or environments. The historical description information field is used to act as the database to which the repeated description information field belongs to obtain the repeated description information field. The historical description information field can be a data set.
[0099] In an optional embodiment, in the step of obtaining the repeated description information field from the historical description information field, the exactly same historical description information field can be used as the overlapping description information field, or it can be achieved by calculating the similarity of the historical description information field. Exemplarily, a word segmentation algorithm can be used. By means of content acquisition methods such as character recognition, the historical description information fields with different expressions are analyzed to obtain the same description information field, and the description information fields with the same semantics are used as the repeated description information fields; alternatively, any historical description information field can be selected as the standard field, and the similarity between other historical description information fields and the standard field is calculated through a machine learning model or the like. If the similarity is within the same range, it is determined as the repeated description information field.
[0100] In an optional embodiment, when auditing the repeated description information field, the repeated project information field that passes the audit is used as the first description information field. When auditing the repeated description information field, any dimension can be selected, including the audit of the file storage channel, the audit of the file acquisition channel, and the overall audit of the file acquisition, etc. Exemplarily, in an optional embodiment, for the confirmation interaction between the information receiver and the information provider, when the information receiver of the requirement file sends a request to generate the first description information field, the information provider who needs to extract the file conducts an audit and confirmation to judge whether the first description information field meets what the information provider can provide; in an optional embodiment, professional personnel such as interface engineers need to conduct an audit and confirmation in terms of technology to judge whether the first description information field can be implemented and whether it will affect other fields, etc., for functional audit; in an optional embodiment, general managers and other comprehensive management personnel need to conduct an audit to judge whether there is repetitive work and whether it will affect the project, etc., for overall audit.
[0101] In this embodiment, by aggregating, inducing and extracting complex and scattered data, the data complexity is reduced. By generating the first description information field, the required number of templates is reduced, making the first file acquisition template more suitable for various fields or projects, and the number of template identifiers to be input is reduced. And the project information field that passes the audit can better meet the project requirements and control the file acquisition process with high precision.
[0102] In one embodiment, such as Figure 6As shown, the steps of obtaining the file management information by filling the target template with the application file scheduling time, project identifier, and file encoding include:
[0103] Step 602, based on the fields of the target template, obtain the project description information corresponding to the project identifier, and obtain the file description information to be extracted corresponding to the file encoding.
[0104] In an optional embodiment, for the convenience of each project to schedule the corresponding files to be extracted, the fields of the target template include: at least two project description information fields and file description information fields to be extracted, which are used to be filled with the corresponding information descriptions or descriptive information. Information description can also be called information resource description, which refers to the activity of analyzing, selecting, and recording the subject content, formal features, physical forms, etc. of information resources according to the needs of information organization and retrieval.
[0105] In an optional embodiment, use the fields of the target template and the project identifier as an index, and select the target project description information from the project description information corresponding to the project identifier to obtain the project description information corresponding to the project identifier; or, use the fields of the target template and the file encoding as an index, and select the target file description information from the file description information corresponding to the file encoding to obtain the file description information corresponding to the file encoding.
[0106] Step 604, based on the mapping relationship between the file scheduling time and the category of the file encoding, determine the file scheduling time of the file to be extracted.
[0107] In an optional embodiment, file encodings belonging to the same category will have the same file scheduling time. This same file scheduling time can refer to having the same time length, or can be scheduled according to the same time point. If scheduled according to the same time point, it can be scheduled between two time points or between two time periods.
[0108] In an optional embodiment, if there are multiple files to be extracted, the files belonging to the same category have a category identifier. After matching through this category identifier, synchronously assign the file scheduling time to the files to be extracted with the same category identifier to establish the file scheduling time of each file to be extracted.
[0109] Step 606, fill the target template with the project description information, the file description information to be extracted, and the file scheduling time of the file to be extracted, and construct the corresponding relationship between the file scheduling time, the project, and the file to be extracted.
[0110] In an optional embodiment, the correspondence relationship among the file scheduling time, the project, and the file to be extracted can be a non-mapping relationship. For example, when the file scheduling time is missing, the correspondence relationship between the project and the file to be extracted can still be established, or a default file scheduling time can be set to establish the correspondence relationship among the default time, the project, and the file to be extracted.
[0111] In an optional embodiment, the correspondence relationship among the file scheduling time, the project, and the file to be extracted can be a mapping relationship, and this mapping relationship can be one-to-one or one-to-many. For example, at the same file scheduling time, multiple projects can obtain multiple files.
[0112] In this embodiment, the correspondence relationship between the description information is established and, in conjunction with the file scheduling time, file management information that is convenient for retrieval is formed. The file management information includes various types of description information, facilitating batch scheduling of big data and extracting multiple files for multiple projects simultaneously.
[0113] In one embodiment, after mapping and filling the template, there may be situations such as data omission or overflow. For example, when the data volume after filling a certain field is too large and exceeds the threshold specified by the template, the data may not be filled in. Based on this, steps for correcting the data are also required. As Figure 7 shown, after filling the target template with the application project description information, the file description information to be extracted, and the file scheduling time of the file to be extracted, it includes:
[0114] Step 702, obtain the original information corresponding to the fields of the target template from the project description information and the file description information.
[0115] In an optional embodiment, using the fields of the target template as an index, search for relevant data from the project description information and the file description information to form the original information. The original information can be specific fields, or the identifiers corresponding to the specific fields, or specific data ranges.
[0116] In an optional embodiment, after obtaining the original information corresponding to the fields of the target template, a data set is formed, and duplicate information in this data set needs to be removed to obtain a mapping table with a mapping relationship or other data sets without duplicate information.
[0117] Step 704, compare the matched original information with the file management information. If the original information is more than the file management information, obtain the difference information between the original information and the file management information.
[0118] In an optional embodiment, comparing the matching original information with the file management information may involve comparing information with the same identifier. If the information has the same identifier, it is determined to be a match. In an optional embodiment, the similarity of the information after semantic extraction may be compared. When the information similarity reaches a threshold, the original information and the file management information are determined to be a match.
[0119] In an optional embodiment, if the original information is more than the file management information, the difference information between the original information and the file management information is obtained, including: determining whether the total amount of the original information is more than the file management information; or, determining whether the specific fields in the original information are more than the specific fields in the file management information; or, determining whether the data length in the original information is more than the data length in the file management information.
[0120] In an optional embodiment, the difference information corresponds to specific description information fields. For example, if the project description information in the original information is more than the file management information, the project description difference information can be obtained; if the file description information in the original information is more than the file management information, the file description difference information can be obtained.
[0121] Step 706, supplement the difference information to the file management information to obtain the corrected file management information. The corrected file management information includes the corrected project description information, the corrected file to be extracted, and / or the corrected correspondence.
[0122] In an optional embodiment, different difference information is supplemented to the fields of different target templates. For example, if the project description information in the original information is more than the file management information, the project description difference information can be obtained, and the project description difference information is used to correct the project description information.
[0123] In this embodiment, considering that when filling the template, it will be restricted by factors such as the field length in the template field attributes, which may cause problems such as data loss or data chaos. Therefore, by obtaining the original information corresponding to the fields of the target template to provide evidence and correct the errors in the file management information, refined management is further realized.
[0124] In one embodiment, as Figure 8 shown, filling the target template with the project description information and the file description information to be extracted to obtain the file management information further includes:
[0125] Detecting whether there is a modification to the target template. If there is a modification, based on the modified target template, obtaining the description information corresponding to the project code and / or the file code respectively to obtain the updated description information, and using the updated description information to update the file management information.
[0126] In an optional embodiment, detecting whether the target template has been modified can be achieved by detecting any data in the target template fields or by detecting data related to the target template. For example, it can be detected whether the fields or template identifiers of the target template have changed, or whether the template corresponding to the template identifier in the target template has changed or the corresponding template fields have been changed.
[0127] In an optional embodiment, based on the modified target template, the project description information corresponding to the project code and the file description information to be extracted corresponding to the file code are obtained to obtain updated description information, and at least one data processing is performed. For example, the fields after modifying the target template can be used as an index to re-obtain information, and the corresponding description information can be directly obtained according to the project code and file code to achieve real-time data update. It can also be combined with the steps of data cleaning or other methods of data aggregation and reorganization.
[0128] In this embodiment, by setting a correction mechanism for the target template, only the fields of the target template need to be corrected, and the corresponding files can be automatically updated, realizing large-scale data standardization and refined control.
[0129] In a specific application scenario, the above technology is applied to the relevant background of data extraction. In design institutes in various fields, production data is often utilized and exchanged between different specialties and departments. This action is commonly known as data extraction. For example, if Specialty A needs a certain data from Specialty B, then Specialty B needs to provide relevant materials and parameters to Specialty A. B is the data provider and A is the data recipient. The entire data extraction process requires submitting a request to establish a need, setting a planned completion date, and then the specific responsible designer initiates the data extraction process. Each node has time progress requirements. At the same time, the project office will also track and control the delayed and unclosed needs.
[0130] The entire process of data extraction is implemented in the data extraction system, and online processing is carried out by using information platform application means. Generally, the data extraction system can realize the establishment and maintenance of data extraction requirements. The requirements are divided into different states according to the progress of the current link. When the designer initiates a specific data extraction process, there may be a situation of using the data extraction form multiple times. Version differentiation is adopted, and each version has a receipt opinion, which can monitor and record the entire life cycle of data extraction, and finally perform statistical analysis and query.
[0131] In traditional technologies, generally different data is stored in different databases, and different identities are used to operate the data in different databases to avoid data misoperations and chaos during data provision; or, the data provision is defined based on the data provision types among different specialties, and a corresponding intermediate database is established according to the data provision; taking the intermediate database as a bridge, a multi-specialty data provision and receipt method between two platforms is realized, and subsequent roles read the data provision in the intermediate database and cooperate with the digital model to perform verification, release, distribution, and receipt processing on the data provision. However, for technical fields with large amounts of big data, the data extraction methods of traditional technologies are too rough to control the overall data.
[0132] In a certain environment, if some projects have characteristics such as long construction periods, large amounts of data, and high complexity, for example: projects responsible by design institutes, especially the construction of nuclear power projects, it is necessary to solve the problems of excessive personalized requirements, overly scattered data, duplicate data in different projects, and large workload of data provision requirement management in the data provision system. Standardized, normalized, batch, and templated methods are needed to improve the standardization and efficiency of data, find common ground while reserving differences in complex data for inductive extraction, reduce data complexity, control the number of personalized requirements, and reduce the workload of manual maintenance and management. And the above embodiments of the present application can achieve this effect.
[0133] In a certain technical solution in the above abstract environment, it can be divided into 4 parts: "New Standardized List", "Standardized List Planning", "Standardized List Change", and "Standardized Form Template Maintenance". Each stage can be independent or combined, and all should be protected.
[0134] "New Standardized List": If the server is used as the implementation object, it can correspond to the above "Receiving a file call request, obtaining the file code of the file to be extracted carried by the file call request" and "Selecting the template identifier in the template set according to the file call request, and combining the template fields corresponding to the selected template identifier to form a target template". It is the basis of standardized data provision. The data provision management role can manually add standardized lists one by one through this function, or add data in batches through the EXCEL import method. This list mainly includes the design stage, sub-item, system, data provision specialty, data provision department, item number, receipt specialty, receipt department, key data provision, data name, requirement description, data provision classification, unit number, operation code, and template code; these fields are all standardized general fields and are applicable to all projects.
[0135] "Standardized list planning": If the server is the implementation object, it can correspond to the above "obtaining the file scheduling time mapped to the category of the file code" and "applying the file scheduling time, the project identifier, and the file code to fill the target template to obtain file management information". It can be a process of generating multiple project requirements using a target template.
[0136] In an optional implementation manner, select the list to be planned, and then select the project to be planned. At this time, the intermediate data for planning will be automatically formed. The steps for forming the intermediate data can be:
[0137] There is an operation code in the standardized list. Through the operation code and the specific project, the IED classification can be obtained from the operation planning system. The IED classification can be a file publication plan, and the file publication plan includes Category 1, Category 2, Category 3, and Category 4.
[0138] After the intermediate data is formed, these data will match the planned completion dates for each status according to the operation code and the mapping table, as shown in Figure 8 In an optional embodiment, this step includes:
[0139] According to the IED classification, obtain the FIN activity cycle and / or the FRZ activity cycle respectively based on the IED classification mapping periodic table; among them, the FIN activity cycle is the second activity cycle for obtaining the time of a relatively complete version, and the FRZ activity cycle is the third activity cycle for obtaining the time of the frozen version.
[0140] If the planned time is the initial planned date, subtract the corresponding activity cycles from the initial planned date respectively to obtain the final FIN and FRZ planned dates for each operation code, and then select the two nearest dates as the FIN planned completion date and the FRZ planned completion date for this project requirement; the FIN planned completion date and the FRZ planned completion date are used to inform the professionals that this requirement needs to complete the information submission process and solidification within this time, including the times for both the FIN and FRZ states, and it is also convenient for project managers to conduct progress control and early warning reminders.
[0141] In an optional embodiment, use the filled target template as the selected project to confirm the specific project requirements, and the system will automatically supplement and complete other fields, including the information submission and receiving responsible persons and the planned completion dates for the three states of PRE, FIN, and FRZ. If there are some incomplete data at this time, such as too many information submission and receiving responsible persons and the planned dates not being matched, etc., then the information submission management role or some programs need to correct and fill the planned dates. When the verification information passes, the project can be imported to achieve a complete requirement.
[0142] In an optional embodiment, the project requirements are the actual effective information submission requirements, with clear projects and PRE / FIN / FRZ planned dates. The requirements planned through the standardized list belong to the standardized requirements, that is, only the project requirements generated by filling the first document acquisition template are the standardized requirements, while the project requirements generated in cooperation with the above-mentioned second document acquisition template belong to the general requirements.
[0143] "Standardized list change": It corresponds to the step of "forming the first document acquisition template". For the newly added list data, there will also be maintenance and modification situations in the future, especially the feedback from downstream professional designers. Because the standardized list has a greater impact, strict control must be carried out, so it is necessary to go through the change application process for approval, which includes multi-level approvals, specifically as Figure 9 shown.
[0144] "Standardized form template maintenance": Corresponding to "obtaining the target template set", the form template can be added, modified, or deleted. The entity content is a standard editing template such as a Word document. When a new item is added to the standardized list, a template can be selected for association. If the list is associated with a template, when the project requirements initiated are used to draft the information submission form, a template will be automatically copied as an attachment to the information submission form process, and online editing can be achieved, which is very convenient for designers to fill in the corresponding data, realizing the automatic association and application of the template, eliminating the need for users to upload the template again, and achieving unified standardization.
[0145] In this scenario, an implementation plan for standardized management of an information submission system applicable to design institutes is provided, which has great advantages in terms of cost or efficiency in aspects such as generating project requirements and integration.
[0146] In terms of generating project requirements and cost, based on a target template, different project requirements planned have a one-to-many relationship. The project requirements generated from the same template can be standardized data, and subsequent designers do not need to fill in these standardized fields again, and the system will automatically bring them out. This standardizes the diversity of information submission data and reduces the complexity of requirements. Each project requirement can selectively set three planned dates of PRE / FIN / FRZ to achieve the setting, monitoring, and early warning of the time schedule. The requirements planned through the standardized list, because they have the basic data of the same set of templates, once the basic data is modified, all associated project requirements will be automatically batch-modified, and it can be used for traceability to view the modification and approval records, reducing the management and maintenance cost of documents. In short, by using the method of batch planning projects through the standardized list to establish project requirements, standardization, normativity, and unity are achieved, and the workload of establishing requirements is also reduced.
[0147] For integration, modular development is adopted, which has no impact on the existing information submission system and can achieve seamless connection of new and old functions and new and old data. Moreover, by filling in the template code when adding a new list, when the code filled in the list is the same as the code in the list filling and template maintenance module, an association is generated. Through the project requirements planned by this list, the corresponding maintenance template information will be automatically captured and carried when applied later.
[0148] To more conveniently understand the technical solution of this application, it is discussed from the perspective of user operations, including the establishment method of standardized information submission, the construction method of standardized lists, and the modification method of standardized lists: Among them:
[0149] The establishment method of standardized information submission includes:
[0150] Summarize and streamline the general attributes and requirements of ordinary personalized needs into a standardized list. After creating a new list, select the corresponding project to batch plan project requirements. The system automatically calculates the completion time of each status according to the association of the operation code and the time mapping table. After passing the verification, specific information submission requirements are automatically generated. Among them, the general attributes and requirements specifically refer to the fields when adding a new standardized list. When planning different projects later, these fields are common and fixed, belonging to template data.
[0151] The construction method of standardized lists includes:
[0152] First, establish a module for template maintenance to implement basic database functions such as addition, deletion, modification, and query. In the standardized list, one of all templates can be selected for association. After establishing the relationship, when the specific requirements related to this standardized list go through the information submission form process, the system will automatically copy a copy as a template attachment and can achieve online editing.
[0153] The modification method of standardized lists includes:
[0154] Conduct process approval and record for the modification of the standardized list. After the review is passed, the specific executor will perform the modification operation, and the node before execution can be rolled back.
[0155] It should be understood that although Figure 2-7 the steps in the flowchart are shown in sequence according to the arrows, these steps are not necessarily executed in the order indicated by the arrows. Unless there is a clear description in this article, the execution of these steps has no strict order limit, and these steps can be executed in other orders. Moreover, Figure 2-7At least some of the steps may include multiple steps or multiple stages, and these steps or stages do not necessarily need to be executed and completed at the same moment. Instead, they can be executed at different moments, and the execution order of these steps or stages does not necessarily need to be sequential. Instead, they can be executed alternately or in rotation with at least some of the steps or stages in other steps or other steps.
[0156] In one embodiment, as Figure 10 shown, a file acquisition device is provided, including: an acquisition time determination module, a template acquisition module, a correspondence determination module, and a file extraction module, where:
[0157] The acquisition time determination module is configured to receive a file call request, obtain the file code and project identifier of the file to be extracted carried by the file call request, obtain the category to which the file code belongs, and obtain the file scheduling time mapped to the category to which the file code belongs;
[0158] The template acquisition module is configured to obtain a target template set based on the file call request, select a template identifier in the template set according to the file call request, and combine the template fields corresponding to the selected template identifier to form a target template;
[0159] The correspondence determination module is configured to fill the target template with the mapped file scheduling time, the project identifier, and the file code to obtain file management information, and the file management information is used to indicate the correspondence between the file scheduling time, the project, and the file to be extracted;
[0160] The file extraction module is configured to, at the file scheduling time, follow the correspondence, and obtain the corresponding file to be extracted according to the project.
[0161] In one embodiment, the time determination module includes a category determination unit, a period determination unit, an initial time acquisition unit, and a scheduling time calculation unit, where:
[0162] The category determination unit is configured to obtain a job classification mapping table, and determine the category of the file to be extracted to which the file code belongs according to the job classification mapping table.
[0163] The period determination unit is configured to obtain a period mapping table corresponding to the category of the file to be extracted, and estimate the activity period corresponding to the category of the file to be extracted based on the period mapping table.
[0164] The initial time acquisition unit is configured to obtain the initial time corresponding to the file code, and the initial time is the time when the file call request is received.
[0165] A scheduling time calculation unit, configured to calculate based on the initial time and the estimated activity period to obtain the file scheduling time.
[0166] In one embodiment, the template acquisition module includes a template selection unit and a template combination unit, where:
[0167] The template selection unit is configured to obtain the template code in the file call request and match the template code with the template identifiers in the template set;
[0168] The template combination unit is configured to, when the matching is successful, obtain the template fields corresponding to the matched template identifier from the template set according to the matched template identifier, and combine the selected template fields to obtain the target template.
[0169] In one embodiment, the template combination unit includes a first template subunit, a template association subunit, a second template subunit, and a template construction subunit, where:
[0170] The first template subunit is configured to obtain a first file acquisition template, where the first file acquisition template includes a first description information field and an associated template identifier field;
[0171] The template association subunit is configured to fill the matched associated template identifier into the associated template identifier field to obtain a target associated identifier;
[0172] The second template subunit is configured to obtain a second file acquisition template corresponding to the target associated identifier, where the second file acquisition template includes a second description information field;
[0173] The template construction subunit is configured to add the second description information field to the first file acquisition template to generate the target template.
[0174] In one embodiment, the correspondence determination module includes a file information acquisition unit, a scheduling time determination unit, and a correspondence construction unit, where:
[0175] The file information acquisition unit is configured to obtain the project description information corresponding to the project identifier and the file description information to be extracted corresponding to the file code based on the fields of the target template;
[0176] The scheduling time determination unit is configured to determine the file scheduling time of the file to be extracted based on the mapping relationship between the file scheduling time and the category to which the file code belongs;
[0177] A correspondence building unit, configured to fill the target template with the project description information, the file description information to be extracted, and the file scheduling time of the file to be extracted, and build the correspondence among the file scheduling time, the project, and the file to be extracted.
[0178] In one embodiment, the correspondence building unit includes a correction information acquisition subunit, a difference information acquisition subunit, and a management information correction subunit, where:
[0179] The correction information acquisition subunit is configured to acquire the original information corresponding to the fields of the target template from the project description information and the file description information;
[0180] The difference information acquisition subunit is configured to compare the matched original information with the file management information. If the original information is more than the file management information, acquire the difference information between the original information and the file management information;
[0181] The management information correction subunit is configured to supplement the difference information to the file management information to obtain the corrected file management information, where the corrected file management information includes the corrected project description information, the corrected file to be extracted, and / or the corrected correspondence.
[0182] In one embodiment, the correspondence determination module further includes a template information reconstruction unit. The template information reconstruction unit is configured to detect whether there is any modification to the target template. If there is a modification, based on the modified target template, acquire the description information corresponding to the project code and / or the file code respectively to obtain the updated description information, and use the updated description information to update the file management information.
[0183] For the specific limitations of the file acquisition device, reference may be made to the limitations on the file acquisition method in the foregoing text, which will not be elaborated here. Each module in the foregoing file acquisition device may be implemented in whole or in part by software, hardware, and their combination. The foregoing modules may be embedded in the processor in the computer device in hardware form or be independent of it, or may be stored in the memory in the computer device in software form, so that the processor can call and execute the operations corresponding to the foregoing modules.
[0184] In one embodiment, a computer device is provided. The computer device may be a server, and its internal structure diagram may be as Figure 11As shown in the figure. The computer device includes a processor, a memory, and a network interface connected by a system bus. Among them, the processor of the computer device is used to provide computing and control capabilities. The memory of the computer device includes a non-volatile storage medium and an internal memory. The non-volatile storage medium stores an operating system, a computer program, and a database. The internal memory provides an environment for the operation of the operating system and the computer program in the non-volatile storage medium. The database of the computer device is used to store file acquisition data. The network interface of the computer device is used to communicate with an external terminal through a network connection. When the computer program is executed by the processor, it implements a file acquisition method.
[0185] Those skilled in the art can understand that Figure 11 the structure shown in the figure is only a block diagram of some structures related to the solution of the present application, and does not constitute a limitation on the computer device to which the solution of the present application is applied. The specific computer device may include more or fewer components than those shown in the figure, or combine some components, or have different component arrangements.
[0186] In one embodiment, a computer device is further provided, including a memory and a processor. A computer program is stored in the memory. When the processor executes the computer program, the steps in the above method embodiments are implemented.
[0187] In one embodiment, a computer-readable storage medium is provided, on which a computer program is stored. When the computer program is executed by the processor, the steps in the above method embodiments are implemented.
[0188] Those of ordinary skill in the art can understand that all or part of the processes of implementing the above method embodiments can be completed by instructing relevant hardware through a computer program. The computer program can be stored in a non-volatile computer-readable storage medium. When the computer program is executed, it can include the processes of the above method embodiments. Among them, any reference to a memory, storage, database, or other medium used in the various embodiments provided in the present application can include at least one of non-volatile and volatile memories. Non-volatile memory can include read-only memory (ROM), magnetic tape, floppy disk, flash memory, or optical memory, etc. Volatile memory can include random access memory (RAM) or external cache memory. By way of illustration and not limitation, RAM can be in various forms, such as static random access memory (SRAM) or dynamic random access memory (DRAM), etc.
[0189] The technical features of the above embodiments can be combined arbitrarily. For the sake of concise description, not all possible combinations of the technical features in the above embodiments are described. However, as long as there is no contradiction in the combination of these technical features, it should be considered as the scope described in this specification.
[0190] The above-described embodiments merely represent several implementation manners of the present application. The description is relatively specific and detailed, but it should not be construed as a limitation on the scope of the invention patent. It should be noted that for those of ordinary skill in the art, without departing from the concept of the present application, several modifications and improvements can still be made, and these all belong to the protection scope of the present application. Therefore, the protection scope of the patent of the present application shall be subject to the appended claims.
Claims
1. A file acquisition method, characterized in that, The method includes: Receiving a file call request, obtaining the file code and project identifier of the file to be extracted carried in the file call request, obtaining the category to which the file code belongs, and obtaining the file scheduling time mapped to the category to which the file code belongs; Obtaining a target template set based on the file call request, selecting a template identifier from the template set according to the file call request, and combining the template fields corresponding to the selected template identifier to form a target template; Populating the target template with the file scheduling time, the project identifier, and the file code to obtain file management information, where the file management information is used to indicate the correspondence between the file scheduling time, the project identifier, and the file to be extracted; At the file scheduling time, following the correspondence, obtaining the corresponding file to be extracted according to the project identifier; Among them, populating the target template with the file scheduling time, the project identifier, and the file code to obtain file management information includes: Based on the fields of the target template, obtaining the project description information corresponding to the project identifier, and obtaining the description information of the file to be extracted corresponding to the file code; determining the file scheduling time of the file to be extracted based on the mapping relationship between the file scheduling time and the category to which the file code belongs; populating the target template with the project description information, the description information of the file to be extracted, and the file scheduling time of the file to be extracted, and constructing the correspondence between the file scheduling time, the project, and the file to be extracted.
2. The method according to claim 1, characterized in that, The obtaining the category to which the file code belongs and obtaining the file scheduling time mapped to the category to which the file code belongs includes: Obtaining a job classification mapping table, and determining the category of the file to be extracted to which the file code belongs according to the job classification mapping table; Obtaining a cycle mapping table corresponding to the category of the file to be extracted, and estimating the activity cycle corresponding to the category of the file to be extracted based on the cycle mapping table; Obtaining the initial time corresponding to the file code, where the initial time is the time when the file call request is received; Calculating based on the initial time and the estimated activity cycle to obtain the file scheduling time.
3. The method according to claim 1, characterized in that The selecting a template identifier from the template set according to the file call request and combining the template fields corresponding to the selected template identifier to form a target template includes: Obtaining the template code in the file call request, and matching the template code with the template identifiers in the template set; If the match is successful, then according to the matched template identifier, obtaining the template fields corresponding to the matched template identifier from the template set, and combining the selected template fields to obtain the target template.
4. The method according to claim 3, wherein The template identifier includes an associated template identifier, and the obtaining the template fields corresponding to the matched template identifier from the template set according to the matched template identifier and combining the selected template fields to obtain the target template includes: Obtaining a first file acquisition template, where the first file acquisition template includes a first description information field and an associated template identifier field; Populating the associated template identifier field with the matched associated template identifier to obtain a target associated identifier; Obtain a second file acquisition template corresponding to the target association identifier, where the second file acquisition template includes a second description information field; Add the second description information field to the first file acquisition template to generate the target template.
5. The method according to claim 1, wherein After filling the target template with the project description information, the file description information to be extracted, and the file scheduling time of the file to be extracted, it includes: Obtain the original information corresponding to the fields of the target template from the project description information and the file description information; Compare the matched original information with the file management information. If the original information is more than the file management information, obtain the difference information between the original information and the file management information; Supplement the difference information to the file management information to obtain the corrected file management information, where the corrected file management information includes the corrected project description information, the corrected file to be extracted, and / or the corrected correspondence.
6. The method according to any one of claims 1 to 5, characterized in that After filling the target template with the file scheduling time, the project identifier, and the file code to obtain the file management information, it further includes: Detect whether there is a modification to the target template. If there is a modification, based on the modified target template, obtain the description information corresponding to the project code and / or the file code respectively to obtain the updated description information, and use the updated description information to update the file management information.
7. A file acquisition device, characterized in that, The device includes: An acquisition time determination module, configured to receive a file call request, obtain the file code and the project identifier of the file to be extracted carried in the file call request, obtain the category to which the file code belongs, and obtain the file scheduling time mapped to by the category to which the file code belongs; A template acquisition module, configured to obtain a target template set based on the file call request, select a template identifier in the template set according to the file call request, and combine the template fields corresponding to the selected template identifier to form a target template; A correspondence determination module, configured to fill the target template with the file scheduling time, the project identifier, and the file code to obtain file management information, where the file management information is used to indicate the correspondence between the file scheduling time, the project identifier, and the file to be extracted; A file extraction module, configured to obtain the corresponding file to be extracted according to the project identifier at the file scheduling time following the correspondence; Among them, the correspondence determination module includes a file information acquisition unit, a scheduling time determination unit, and a correspondence construction unit; The file information acquisition unit is configured to obtain the project description information corresponding to the project identifier and the file description information to be extracted corresponding to the file code based on the fields of the target template; The scheduling time determination unit is configured to determine the file scheduling time of the file to be extracted based on the mapping relationship between the file scheduling time and the category to which the file code belongs; The correspondence construction unit is configured to fill the target template with the project description information, the file description information to be extracted, and the file scheduling time of the file to be extracted, and construct the correspondence between the file scheduling time, the project, and the file to be extracted.
8. The device according to claim 7, characterized in that The time determination module includes a category determination unit, a period determination unit, an initial time acquisition unit, and a scheduling time calculation unit, where: The category determination unit is configured to obtain a job classification mapping table, and determine the category of the file to be extracted to which the file encoding belongs according to the job classification mapping table; The period determination unit is configured to obtain a period mapping table corresponding to the category of the file to be extracted, and estimate the activity period corresponding to the category of the file to be extracted based on the period mapping table; The initial time acquisition unit is configured to obtain the initial time corresponding to the file encoding, and the initial time is the time when the file call request is received; The scheduling time calculation unit is configured to perform calculations based on the initial time and the estimated activity period to obtain the file scheduling time.
9. A computer device, comprising a memory and a processor, the memory storing a computer program, characterized in that, When the processor executes the computer program, the steps of the method described in any one of claims 1 to 6 are implemented.
10. A computer-readable storage medium having a computer program stored thereon, characterized in that, When the computer program is executed by the processor, the steps of the method described in any one of claims 1 to 6 are implemented.
Citation Information
Patent Citations
Method and system for realizing automatic management of template files
CN106933598A
Data processing method and device, server and computer readable storage medium
CN112148509A