A method, device, equipment and medium for summarizing valuation tables
By preconfiguring the matching algorithm of table headers and subject keywords, the problem of low summarization efficiency of valuation tables in different formats is solved, and automated valuation table summary is realized, reducing manual intervention and improving summary efficiency.
Patent Information
- Application Number
- CN202111261885.7
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2021-10-28
- Publication Date
- 2025-06-24
- Estimated Expiration
- 2041-10-28
AI Technical Summary
In the prior art, since the valuation table formats used by companies in different asset industries are different, the names and codes of the same main account are inconsistent in different valuation tables, which increases the difficulty and workload of summarizing the valuation tables and reduces the summary efficiency.
By preconfiguring the corresponding relationship between the header keywords, the subject keywords and the subjects, and the corresponding relationship between the subjects and codes, the matching algorithm is used to determine the header and content data in the valuation table, and update it according to the standard account names and code sequences to realize automatic summary of valuation tables in different formats.
The need to manually configure templates or rules for each format reduces the workload of staff, improves the efficiency of valuation table summary, and ensures the accuracy and consistency of summary.
Smart Images

Figure CN113935295B_ABST
Abstract
Description
Technical Field
[0001] The present invention relates to the technical field of data processing, and particularly to a method, device, equipment and medium for summarizing valuation tables. Background Art
[0002] Since the formats of the valuation tables adopted by existing major asset industry companies are different. For example, asset management industry companies use different valuation systems, or different versions of the same valuation system, or the same version but different valuation configurations, etc., which will all result in the names and codes of the same main subject being different in different valuation tables, increasing the difficulty of summarizing different valuation tables.
[0003] In order to summarize different valuation tables, generally, the rules or templates for summarizing each valuation table are manually pre-configured and saved in the device, and then the device can summarize each valuation table according to the rules or templates of each valuation table. For this method, since it is necessary to manually set the corresponding rules or templates for different types of valuation tables, the workload of manual labor is increased, and the efficiency of summarizing each valuation table is also reduced due to the problem of manual efficiency.
[0004] Therefore, how to quickly and accurately summarize different formats of valuation tables has become an urgent technical problem to be solved. Summary of the Invention
[0005] Embodiments of the present invention provide a method, device, equipment and medium for summarizing valuation tables, so as to solve the problem of low efficiency in summarizing different formats of valuation tables in the prior art.
[0006] Embodiments of the present invention provide a method for summarizing valuation tables, and the method includes:
[0007] Match the data in the valuation table to be summarized with preset header keywords to determine each header included in the data and the content data of each header in the data;
[0008] Match the target content data with the header of the subject name with preset subject keywords to determine the target subject keywords that match the target content data;
[0009] Determine the standard subject name corresponding to the target content data according to the subject name of the target subject corresponding to the target subject keywords, and determine the standard subject code sequence corresponding to the target content data according to the code of the target subject;
[0010] Update the target content data according to the standard subject name, and update the content data corresponding to the target content data in the content data with the header of the subject code according to the standard subject code sequence;
[0011] Determine the summarized valuation table based on the respective header tables and the content data of the respective header tables.
[0012] Furthermore, the subject keyword includes an item keyword and an item attribute keyword. Obtaining the subject name of the target subject corresponding to the target subject keyword includes:
[0013] Match the target content data with the item keyword and the item attribute keyword respectively to determine the target item keyword that matches the target content data and the target item attribute keyword that matches the target content data;
[0014] Determine the subject name of the target subject based on the target item name of the item corresponding to the target item keyword and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword.
[0015] Furthermore, determining the subject name of the target subject based on the target item name of the item corresponding to the target item keyword and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword includes:
[0016] If there are at least two target item keywords and at least two target item attribute keywords, determine the item priority of the item corresponding to each target item keyword according to the preset item priority, and determine the attribute priority of the item attribute classification corresponding to each target item attribute keyword according to the priority of the item attribute classification included in the preset target item;
[0017] Concatenate the respective target item names according to each item priority to obtain a first concatenated sequence; and
[0018] Concatenate the respective target item attribute classification names according to each attribute priority to obtain a second concatenated sequence;
[0019] Determine the subject name of the target subject based on the first concatenated sequence and the second concatenated sequence.
[0020] Furthermore, obtaining the code corresponding to the target subject includes:
[0021] Concatenate the codes corresponding to the respective target item names according to each item priority to obtain a first code sequence; and
[0022] According to each of the attribute priorities, splice the codes corresponding to the classification names of each of the target entry attributes to obtain a second code sequence;
[0023] Determine the code corresponding to the target subject according to the first code sequence and the second code sequence.
[0024] Further, the determining the summarized valuation table according to each of the table headers and the content data of each of the table headers includes:
[0025] Match the target content data with preset place keyword terms to determine the place keyword terms that match the target content data;
[0026] Determine the target trading venue corresponding to the matched place keyword terms;
[0027] Determine the summarized valuation table according to each of the table headers, the content data of each of the table headers, and the target trading venue corresponding to the target content data.
[0028] Further, the method further includes:
[0029] If the summarized valuation table contains an asymmetric total item, verify the data format of the first value in the asymmetric total item; wherein, the asymmetric total item is a total item in the summarized valuation table that is not a symmetric total item, and the symmetric total item has a corresponding table header in the summarized valuation table.
[0030] Further, the verifying the data format of the first value in the asymmetric total item includes:
[0031] Obtain the valuation table summarized before the summarized valuation table; obtain the second value in the asymmetric total item from the obtained valuation table; determine whether the data format of the second value is the same as the data format of the first value; or
[0032] Determine whether the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item.
[0033] An embodiment of the present invention provides a valuation table summarization device, and the device includes:
[0034] A first processing unit, configured to match the data in the valuation table to be summarized with preset table header keyword terms to determine each of the table headers included in the data and the content data of each of the table headers in the data;
[0035] A second processing unit, configured to match the target content data with a header of subject name against preset subject keyword terms, and determine target subject keyword terms that match the target content data;
[0036] A third processing unit, configured to determine a standard subject name corresponding to the target content data according to the subject name of the target subject corresponding to the target subject keyword terms, and determine a standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject;
[0037] An update unit, configured to update the target content data according to the standard subject name, and update the content data corresponding to the target content data in the content data with a header of subject code according to the standard subject code sequence;
[0038] A summarization unit, configured to determine a summarized valuation table according to the respective headers and the content data of the respective headers.
[0039] Further, the third processing unit is specifically configured to, if the subject keyword terms include item keyword terms and item attribute keyword terms, match the target content data with the item keyword terms and the item attribute keyword terms respectively, and determine target item keyword terms that match the target content data, and target item attribute keyword terms that match the target content data; and determine the subject name of the target subject according to the target item name of the item corresponding to the target item keyword terms and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword terms.
[0040] Further, the third processing unit is specifically configured to, if there are at least two target item keyword terms and at least two target item attribute keyword terms, determine the item priority of the item corresponding to each target item keyword term according to a preset item priority, and determine the attribute priority of the item attribute classification corresponding to each target item attribute keyword term according to the priority of the item attribute classification included in the preset target item; splice the respective target item names according to each item priority to obtain a first splicing sequence; and splice the respective target item attribute classification names according to each attribute priority to obtain a second splicing sequence; and determine the subject name of the target subject according to the first splicing sequence and the second splicing sequence.
[0041] Further, the third processing unit is specifically configured to splice the codes corresponding to each of the target entry names according to each entry priority to obtain a first code sequence; and splice the codes corresponding to each of the target entry attribute classification names according to each attribute priority to obtain a second code sequence; and determine the code corresponding to the target subject according to the first code sequence and the second code sequence.
[0042] Further, the device further includes: a fourth processing unit;
[0043] The fourth processing unit is configured to match the target content data with preset place keywords to determine the place keywords matching the target content data; and determine the target trading place corresponding to the matching place keywords;
[0044] The summarizing unit is specifically configured to determine the summarized valuation table according to the respective table headers, the content data of each table header, and the target trading place corresponding to the target content data.
[0045] Further, the device further includes: a verification unit;
[0046] The verification unit is configured to, if the summarized valuation table includes an asymmetric total item, verify the data format of the first value in the asymmetric total item; wherein, the asymmetric total item is a total item in the summarized valuation table that is not a symmetric total item, and the symmetric total item has a corresponding table header in the summarized valuation table.
[0047] Further, the verification unit is specifically configured to obtain the valuation table summarized before the summarized valuation table; obtain the second value in the asymmetric total item from the obtained valuation table; determine whether the data format of the second value is the same as the data format of the first value; or determine whether the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item.
[0048] An embodiment of the present invention provides an electronic device, where the electronic device includes a processor, and the processor is configured to implement the steps of any one of the above valuation table summarization methods when executing a computer program stored in a memory.
[0049] An embodiment of the present invention provides a computer-readable storage medium, which stores a computer program, and the computer program implements the steps of any one of the above valuation table summarization methods when executed by a processor.
[0050] Since the header keyword terms, subject keyword terms, the correspondence between subject keyword terms and subjects, and the correspondence between subjects and codes are pre-configured, during the process of summarizing the valuation tables, by matching the data in the valuation tables to be summarized with the pre-set header keyword terms, it is possible to determine each header included in the valuation tables to be summarized and the content data of each header in the valuation tables to be summarized. Then, the target content data with the header being the subject name is matched with the pre-set subject keyword terms to determine the target subject keyword terms that match the target content data. According to the subject name of the target subject corresponding to the target subject keyword terms, the standard subject name corresponding to the target content data can be determined, avoiding the problem that the subject names of the same subject in valuation tables of different formats are inconsistent, resulting in the inability to summarize valuation tables of different formats. And according to the correspondence between subjects and codes, the code corresponding to the target subject can be obtained, so that according to the code corresponding to the target subject, the standard subject code sequence corresponding to the target content data can be determined, also avoiding the problem that the subject codes of the same subject in valuation tables of different formats are inconsistent, resulting in the inability to summarize valuation tables of different formats. Subsequently, the target content data is updated according to the standard subject name, and the content data corresponding to the target content data in the content data with the header being the subject code is updated according to the standard subject code sequence. Then, according to each header and the content data of each header, the summarized valuation table is determined. Through the above method, valuation tables of different formats can be accurately summarized without the need for staff to pre-set a set of templates or a set of rules for each format of valuation table, reducing the workload of the staff and avoiding the impact of the efficiency of staff in configuring templates or rules on the efficiency of summarizing valuation tables, thereby improving the efficiency of summarizing valuation tables of different formats. BRIEF DESCRIPTION OF THE DRAWINGS
[0051] In order to more clearly illustrate the technical solutions in the embodiments of the present invention, the following will briefly introduce the accompanying drawings required for the description of the embodiments. Obviously, the accompanying drawings in the following description are only some embodiments of the present invention. For those of ordinary skill in the art, other drawings can be obtained based on these drawings without creative efforts.
[0052] Figure 1 FIG.
[0053] Figure 2 FIG.
[0054] Figure 3 FIG.
[0055] Figure 4 Schematic structural diagram of a valuation table summarization device provided by an embodiment of the present invention;
[0056] Figure 5 Schematic structural diagram of an electronic device provided by an embodiment of the present invention. Detailed implementation manners
[0057] The present invention will be further described in detail below with reference to the accompanying drawings. Obviously, the described embodiments are only a part of the embodiments of the present invention, rather than all of them. All other embodiments obtained by those of ordinary skill in the art based on the embodiments of the present invention without creative efforts belong to the scope of protection of the present invention.
[0058] With the implementation of the new regulations on bank wealth management products, regulatory requirements demand penetration management of entrusted assets, and products need to be managed on a net asset value basis. However, there are numerous entrusted institutions in the market, and the formats of the generated valuation tables vary, resulting in the same subject being called different subject names in different valuation tables, and the same subject corresponding to different subject code sequences in different valuation tables, which is not conducive to summarizing valuation tables in different formats. Therefore, how to summarize valuation tables in different formats is very important.
[0059] Currently, different formats of valuation tables can be summarized in the following way: for valuation tables in different formats, templates or rules for summarizing this format of valuation table are set up in advance by manual means, so that the valuation tables in this format can be summarized according to the templates or rules subsequently. For example, when staff set up templates or rules, the various elements that need to be configured include the following content:
[0060] 1) Configure header information. For example, configure columns including nature of the subject, subject code, quantity, unit cost, cost, market value, valuation increment, etc., and information such as whether the format of the cells in the column is text, number, thousand separator, percentage sign, etc.
[0061] 2) Configure subject code information. For example, configure information such as whether there is a separator in the format of the subject code of each subject and what the separator is.
[0062] 3) Configure asset code information. For example, configure rules for obtaining the asset code from the subject code, such as the position of the asset code when there is a separator in the subject code and the position of the asset code when there is no separator in the subject code.
[0063] 4) Configure subject information. For example, configure information such as what character the subject code starts with and the subject corresponding to the subject code.
[0064] 5) Configure trading venue information. For example, configure information such as the trading venue code corresponding to the trading venue information in the content data with the subject name as the table header in the valuation table.
[0065] 6) Configure total item information. For example, configure information such as whether the total items included in the valuation table are symmetric total items, the identification values corresponding to the symmetric total items, and the identification values corresponding to the asymmetric total items.
[0066] 7) Configure valuation table rules. For example, according to the information configured in 1) - 6) above, configure a rule to establish the relationship between this rule and the valuation table, so as to facilitate subsequent summarization of the valuation table in this format according to this rule.
[0067] For this method, the following problems mainly exist:
[0068] First, during the process of configuring templates or rules by staff, there are a very large number of elements to be configured, which requires a high level of professionalism from the staff.
[0069] Second, templates or rules need to be configured for valuation tables in different formats, and the staff need to check the usability of the configured templates or rules. When there are minor changes in a certain format of valuation table, such as the addition of data items or the change of subject codes with unchanged actual meanings, the staff need to reconfigure the templates or rules, resulting in a large amount of repetitive work for the staff during the configuration process, increasing the workload of the staff. Moreover, as the number of valuation tables in different formats increases, it becomes more difficult for the subsequent staff to maintain the pre-configured templates or rules.
[0070] In summary, when using this method, since the staff need to pre-configure the templates or rules corresponding to different formats of valuation tables respectively to ensure that different formats of valuation tables can be summarized subsequently, the staff need a large amount of time to sort out and maintain the pre-configured templates or rules, making the workload of the staff very large. And this method also has relatively high requirements for the professionalism of the staff, and the efficiency of summarizing different formats of valuation tables will be affected by the efficiency of the staff in configuring the templates or rules for this format.
[0071] To solve the problem of low efficiency in summarizing valuation tables in different formats, embodiments of the present invention provide a method, apparatus, device, and medium for summarizing valuation tables. Since the header keyword, subject keyword, the correspondence between the subject keyword and the subject, and the correspondence between the subject and the code are pre-configured, during the process of summarizing the valuation tables, by matching the data in the valuation table to be summarized with the preset header keyword, each header included in the valuation table to be summarized and the content data of each header in the valuation table to be summarized can be determined. Then, the target content data with the header being the subject name is matched with the preset subject keyword to determine the target subject keyword that matches the target content data. According to the subject name of the target subject corresponding to the target subject keyword, the standard subject name corresponding to the target content data can be determined, avoiding the problem that the subject names of the same subject in different formats of valuation tables are inconsistent, resulting in the inability to summarize different formats of valuation tables. And according to the correspondence between the subject and the code, the code corresponding to the target subject can be obtained, so as to determine the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject, also avoiding the problem that the subject codes of the same subject in different formats of valuation tables are inconsistent, resulting in the inability to summarize different formats of valuation tables. Subsequently, the target content data is updated according to the standard subject name, and the content data corresponding to the target content data in the content data with the header being the subject code is updated according to the standard subject code sequence. Then, according to each header and the content data of each header, the summarized valuation table is determined. Through the above method, different formats of valuation tables can be accurately summarized without the need for staff to pre-set a set of templates or a set of rules for each format of valuation table, reducing the workload of the staff and avoiding the impact of the efficiency of staff configuring templates or rules on the efficiency of summarizing valuation tables, improving the efficiency of summarizing different formats of valuation tables.
[0072] Embodiment 1:
[0073] Figure 1 FIG. is a schematic diagram of a process for summarizing a valuation table provided by an embodiment of the present invention, and the process includes:
[0074] S101: Match the data in the valuation table to be summarized with the preset header keyword to determine each header included in the data and the content data of each header in the data.
[0075] The method for summarizing a valuation table provided by an embodiment of the present invention is applied to an electronic device, which can be a smart device, such as a computer, a mobile terminal, etc., or a server, etc.
[0076] When a staff member hopes to summarize a valuation form of a certain format, a summary request for summarizing the valuation form can be input through a smart device, so that the smart device can be controlled to summarize the valuation form through the summary request. Among them, the summary request carries the identifier of the valuation form and the data included in the valuation form.
[0077] It should be noted that there are many specific ways to input the summary request. For example, the way to input the summary request can be to input it by inputting voice information, or to input it by operating the virtual buttons displayed on the display screen of the smart device, etc. In the specific implementation process, it can be flexibly set according to needs and will not be specifically limited here. When the smart device obtains the summary request, it can send the summary request to the electronic device for summarizing the valuation form.
[0078] After the electronic device for summarizing the valuation form receives the summary request, it can parse the summary request to obtain the identifier of the valuation form and the data included in the valuation form carried in the summary request, so as to determine which valuation form is to be summarized currently. Then, the valuation form is determined as the valuation form to be summarized, and based on the valuation form summarization method provided in the embodiments of the present invention, the data in the valuation form to be summarized is processed, so as to realize the summarization of the valuation form to be summarized.
[0079] Since any content data included in the valuation form corresponds to a table header, in order to accurately summarize the data included in the valuation form, the table header keyword corresponding to each table header can be set in advance according to the keyword terms that may be included in the table headers of different formats of valuation forms, such as quantity, net asset value per unit, etc. After obtaining the data of the valuation form to be summarized based on the above embodiments, the data in the valuation form to be summarized is matched with the preset table header keyword terms. Specifically, for each preset table header keyword term, the data in the valuation form to be summarized is matched with the table header keyword term, that is, it is determined whether the data contains the table header keyword term. If it is determined that the data in the valuation form to be summarized matches the table header keyword term, it means that the data contains the table header keyword term, then it is determined that the word in the data that matches the table header keyword term is the table header, and according to the table header name corresponding to the preset table header keyword term, the table header name corresponding to the table header is determined. If it is determined that the data in the valuation form to be summarized does not match the table header keyword term, it means that the data does not contain the table header keyword term, then the next table header keyword term is obtained.
[0080] In a possible implementation manner, the staff member can flexibly process the table header keyword terms corresponding to each pre-configured table header according to needs, such as adding, deleting, replacing, creating new table headers, etc.
[0081] Through the above method, the headers included in the data in the valuation table to be summarized can be determined. Then, according to the columns where the headers are located in the valuation table to be summarized, the content data of each header can be determined.
[0082] For example, the headers included in the data in the valuation table to be summarized include at least one of the following: subject name, subject code, quantity, unit cost, bank deposit, stocks, bonds, funds, taxes and fees, expenses, costs, market value, and valuation increment.
[0083] In a possible implementation manner, for each header, obtain the column where the header is located in the valuation table to be summarized, and determine the data other than the header in this column as the content data of the header.
[0084] S102: Match the target content data with the header of the subject name with the preset subject keyword to determine the target subject keyword that matches the target content data.
[0085] In an actual application scenario, the subject names of the same subject may be different in valuation tables of different formats, resulting in the inability to directly summarize the data in the valuation table based on the subject names in the valuation tables of different formats. Therefore, in order to accurately summarize the valuation table to be summarized, after obtaining the content data of each header based on the above embodiments, the content data with the header of the subject name (for the convenience of description, denoted as target content data) can be processed so that the content data with the header of the subject name corresponds to a unified standard subject name in the summarized valuation table.
[0086] In order to accurately summarize the data included in the valuation table, the subject keywords corresponding to each subject can be preset according to the keywords that the content data with the header of the subject name in the valuation tables of different formats may include, such as bank deposit, demand deposit, etc. After obtaining the target content data based on the above embodiments, for each preset subject keyword, match the target content data with the subject keyword to determine whether the target content data contains the subject keyword. Specifically, if it is determined that the target content data matches the subject keyword, it means that the target content data contains the subject keyword, and then the subject keyword is determined as the target subject keyword. If it is determined that the target content data does not match the subject keyword, it means that the target content data does not contain the subject keyword, and then the next subject keyword is obtained.
[0087] In a possible implementation manner, the staff can flexibly process the subject keywords corresponding to each subject configured in advance, such as adding, deleting, replacing, creating new subjects, etc.
[0088] S103: Determine the standard subject name corresponding to the target content data according to the subject name of the target subject corresponding to the target subject keyword, and determine the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject.
[0089] To facilitate the unification of subject names in valuation tables with different formats, the corresponding relationship between subject keywords and subjects, as well as the corresponding relationship between subjects and subject names, are pre-configured. After determining the target subject keyword based on the above embodiments, according to the pre-configured corresponding relationship between subject keywords and subjects, the subject corresponding to the target subject keyword can be determined (for convenience of description, denoted as the target subject). Then, according to the pre-configured corresponding relationship between subjects and subject names, the subject name corresponding to the target name can be determined. Then, according to the subject name corresponding to the target subject, the standard subject name corresponding to the target content data is determined.
[0090] In a possible implementation manner, if the number of the target subjects is 1, the subject name corresponding to the target subject can be directly determined as the standard subject name corresponding to the target content data.
[0091] In another possible implementation manner, if the number of the target subjects is greater than 1, the subject names corresponding to each target subject can be concatenated, and the concatenated subject name is determined as the standard subject name corresponding to the target content data.
[0092] Similarly, the subject code sequence of the same subject may be different in valuation tables with different formats, resulting in the inability to directly summarize the data in the valuation tables according to the subject code sequences in the valuation tables with different formats. Therefore, to accurately summarize the valuation tables to be summarized, the corresponding relationship between subjects and codes is pre-configured. After determining the target subject based on the above embodiments, according to the pre-configured corresponding relationship between subjects and codes, the code corresponding to the target subject can be determined, and according to the code corresponding to the target subject, the standard subject code sequence corresponding to the target content data is determined.
[0093] In a possible implementation manner, if the number of the target subjects is 1, the code corresponding to the target subject can be directly determined as the standard subject code sequence corresponding to the target content data.
[0094] In another possible implementation manner, if the number of the target subjects is greater than 1, the codes corresponding to each target subject can be concatenated, and the concatenated code is determined as the standard subject code sequence corresponding to the target content data.
[0095] S104: Update the target content data according to the standard subject name, and update the content data corresponding to the target content data in the content data with the subject code as the header according to the standard subject code sequence.
[0096] After determining the standard subject name and the standard code sequence corresponding to the target content data based on the above embodiments, the target content data can be updated according to the standard subject name. And determine the content data corresponding to the target content data in the content data with the subject code as the header. Then, update the determined content data according to the standard subject code sequence, so as to facilitate subsequent summarization of the to-be-summarized valuation table based on each header and the content data of each header.
[0097] S105: Determine the summarized valuation table according to each header and the content data of each header.
[0098] In order to accurately summarize the to-be-summarized valuation table to generate the summarized valuation table, the positions of each header and the content data of each header in the summarized valuation table are pre-configured. After completing the update of the target content data and the update of the content data corresponding to the target content data in the content data with the subject code as the header based on the above embodiments, place each header and the content data included in each current header at the corresponding positions in the valuation table according to the pre-configured positions of each header and its content data in the summarized valuation table, so as to obtain the summarized valuation table.
[0099] Since the artificial only needs to pre-configure the header keyword, subject keyword, the corresponding relationship between the subject keyword and the subject, and the corresponding relationship between the subject and the code once, the update of valuation tables in different formats can be realized, thus reducing the workload of the artificial and the time consumed by the artificial configuration. Moreover, during the process of summarizing the valuation tables, by matching the data in the valuation tables to be summarized with the pre-set header keywords, each header included in the valuation tables to be summarized and the content data of each header in the valuation tables to be summarized can be determined. Then, the target content data with the header of the subject name is matched with the pre-set subject keywords to determine the target subject keywords that match the target content data. According to the subject name of the target subject corresponding to the target subject keywords, the standard subject name corresponding to the target content data can be determined, avoiding the problem that the subject names of the same subject are inconsistent in valuation tables in different formats, resulting in the inability to summarize valuation tables in different formats. And according to the corresponding relationship between the subject and the code, the code corresponding to the target subject can be obtained, so as to determine the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject, also avoiding the problem that the subject codes of the same subject are inconsistent in valuation tables in different formats, resulting in the inability to summarize valuation tables in different formats. Subsequently, the target content data is updated according to the standard subject name, and the content data corresponding to the target content data in the content data with the header of the subject code is updated according to the standard subject code sequence. Then, according to each header and the content data of each header, the summarized valuation table is determined. Through the above method, the valuation tables in different formats can be accurately summarized without the need for staff to pre-set a set of templates or a set of rules for each format of the valuation table, reducing the workload of the staff and avoiding the impact of the efficiency of the staff in configuring the templates or rules on the efficiency of summarizing the valuation tables, improving the efficiency of summarizing valuation tables in different formats.
[0100] Embodiment 2:
[0101] In order to accurately summarize the valuation tables, on the basis of the above embodiments, in the embodiments of the present invention, the subject keywords include item keywords and item attribute keywords, and obtaining the subject name of the target subject corresponding to the target subject keywords includes:
[0102] Matching the target content data with the item keywords and the item attribute keywords respectively to determine the target item keywords that match the target content data and the target item attribute keywords that match the target content data;
[0103] Determine the subject name of the target subject according to the target entry name of the entry corresponding to the target entry keyword and the target entry attribute classification name of the entry attribute classification corresponding to the target entry attribute keyword.
[0104] In the actual application process, since the subject name in the valuation table is generally composed of a primary and secondary entry + at least one entry attribute classification, that is, the subject name is generally in the 1-X mode, that is, a certain 1 primary and secondary entry determines the possible included entry attribute classifications, the number of included entry attribute classifications, and the inclusion relationship between each entry attribute classification. For example, when financial assets are used as the primary and secondary entry, the entry attribute classifications included in this primary and secondary entry include financial instrument classification, investment variety classification, listed and circulated situation classification, and accounting classification. Therefore, in order to accurately determine the standard subject name corresponding to the target content data, in the embodiments of the present invention, the preset subject keywords include preset entry keywords and preset entry attribute keywords. After obtaining the target content data based on the above embodiments, the target content data can be matched with the preset entry keywords and the target content data can be matched with the preset entry attribute keywords. Specifically, for each entry keyword, the target content data is matched with the entry keyword to determine whether the target content data contains the entry keyword. If it is determined that the target content data matches the entry keyword, it means that the target content data contains the entry keyword, and then the entry keyword is determined as the target entry keyword. If it is determined that the target content data does not match the entry keyword, it means that the target content data does not contain the entry keyword, and then the next entry keyword is obtained. At the same time, for each entry attribute keyword, the target content data is matched with the entry attribute keyword to determine whether the target content data contains the entry attribute keyword. Specifically, if it is determined that the target content data matches the entry attribute keyword, it means that the target content data contains the entry attribute keyword, and then the entry attribute keyword is determined as the target entry attribute keyword. If it is determined that the target content data does not match the entry attribute keyword, it means that the target content data does not contain the entry attribute keyword, and then the next entry attribute keyword is obtained.
[0105] After determining the target entry keyword and the target entry attribute keyword based on the above embodiments, according to the pre-configured correspondence between the entry keyword and the entry, determine the entry corresponding to the target entry keyword (for convenience of description, denoted as the target entry), and according to the pre-configured correspondence between the entry attribute keyword and the entry attribute classification, determine the entry attribute classification corresponding to the target entry attribute keyword (for convenience of description, denoted as the target entry attribute classification). According to the entry name of the target entry (for convenience of description, denoted as the target entry name) and the entry attribute classification name of the target entry attribute classification (for convenience of description, denoted as the target entry attribute classification name), determine the subject name of the target subject.
[0106] In a possible implementation manner, when determining the subject name of the target subject according to the target entry name of the entry corresponding to the target entry keyword and the target entry attribute classification name of the entry attribute classification corresponding to the target entry attribute keyword, the following situations mainly include:
[0107] Situation 1: If there is only one target entry keyword and only one target entry attribute keyword, the target entry name and the target entry attribute classification name can be directly concatenated according to the preset name concatenation order, and the concatenated sequence is determined as the subject name of the target subject.
[0108] Case 2: In the actual application process, there is a situation where the subject name contains both a main entry and a secondary entry. For this situation, if the valuation table summarization method provided in the embodiments of the present invention is used to match the subject name with the preset entry keyword, it is very likely that one entry keyword matches the main entry of the subject name, and another entry keyword matches the secondary entry of the subject name, that is, there are two entry keywords that match the subject name. Subsequently, when determining the subject name of the target subject according to the entry names of the entries corresponding to each target entry keyword respectively, there is a problem that it is impossible to determine which entry name corresponding to the target entry keyword comes first and which entry name corresponding to the target entry keyword comes second, that is, it is impossible to accurately determine the positions of the entry names of the entries corresponding to each target entry keyword respectively in the subject name of the target subject, resulting in an inability to accurately determine the subject name of the target subject. Therefore, in order to accurately determine the subject name of the target subject, in the embodiments of the present invention, the priority between each entry is preset. When the target entry and the target entry attribute classification are determined based on the above embodiments, if it is determined that there are at least two target entry keywords and only one target entry attribute keyword, the entry priority of each target entry can be determined according to the preset priority of each entry. According to this entry priority, the names of each target entry are concatenated to determine a concatenation sequence (for the convenience of description, denoted as the first concatenation sequence). According to the preset name concatenation order, the first concatenation sequence and the target entry attribute classification name are concatenated to determine the subject name of the target subject.
[0109] Among them, when pre-configuring the main and secondary entries included in the target subject, for each entry, the entry can be first determined as the main entry, and then all other entries except this entry can be determined as secondary entries. Then, the staff determines according to the actual experience value whether there is any secondary entry that does not coexist with the main entry. If it is determined that a certain secondary entry does not coexist with the main entry, then this secondary entry is not determined as the secondary entry corresponding to the main entry, that is, the main and secondary entries are not determined based on this secondary entry and the main entry.
[0110] Case 3: In the actual application process, there may also be a situation where the subject name contains entry attribute classifications at different levels. For this situation, if the valuation table summarization method provided in the embodiments of the present invention is used to match the subject name with the preset entry attribute keyword, there are likely to be multiple entry keywords matching the subject name. Subsequently, when determining the subject name of the target subject according to each target entry attribute classification name, it is impossible to accurately determine the positions of each target entry attribute classification name in the subject name of the target subject, resulting in the inability to accurately determine the subject name of the target subject. Therefore, in the embodiments of the present invention, a priority is preset among the various entry attribute classifications included in each primary and secondary entry. When the target entry and the target entry attribute classification are determined based on the above embodiments, if it is determined that there is only one target entry keyword and there are at least two target entry attribute keywords, the attribute priority of each target entry attribute classification can be determined according to the preset priority among the various entry attribute classifications included in each primary and secondary entry. According to this attribute priority, the names of the various target entry attribute classifications are concatenated to determine a concatenated sequence (for convenience of description, denoted as the second concatenated sequence). According to the preset name concatenation order, the target entry name and this second concatenated sequence are concatenated to determine the subject name of the target subject.
[0111] Among them, when pre-configuring the entry attribute classifications included in the target subject, for each primary and secondary entry, the various entry attribute classifications that may be included under this primary and secondary entry can be determined first. Then, according to the actual experience value, the staff determines whether there are at least two entry attribute classifications that are not coexistent. If it is determined that there are at least two entry attribute classifications that are not coexistent, then these at least two entry attribute classifications are not determined as the entry attribute classifications included in this primary and secondary subject.
[0112] Case 4: When the target entry and the target entry attribute classification are determined based on the above embodiments, if it is determined that there are at least two target entry keywords and there are at least two target entry attribute keywords, the entry priority of each target entry can be determined according to the preset priority of each entry, and the attribute priority of each target entry attribute classification can be determined according to the preset priority among the various entry attribute classifications included in each primary and secondary entry. Then, according to this entry priority, the names of the various target entries are concatenated to determine a confrontation concatenated sequence, and according to this attribute priority, the names of the various target entry attribute classifications are concatenated to determine the second concatenated sequence. According to the preset name concatenation order, the first concatenated sequence and this second concatenated sequence are concatenated to determine the subject name of the target subject.
[0113] In a possible implementation, while determining the subject name of the target subject, the code corresponding to the target subject can also be determined. The specific determination of the code corresponding to the target subject includes the following situations:
[0114] Situation 1: If there is only one target entry keyword and only one target entry attribute keyword, then the code corresponding to the target entry name and the code corresponding to the target entry attribute classification name can be directly concatenated according to the preset code concatenation order, and the concatenated code sequence is determined as the code corresponding to the target subject.
[0115] Situation 2: If there are at least two target entry keywords and only one target entry attribute keyword, then after determining the entry priority of each target entry, the codes corresponding to the names of each target entry are concatenated according to the entry priority to determine a code sequence (for convenience of description, denoted as the first code sequence). According to the preset code concatenation order, the first code sequence and the code corresponding to the target entry attribute classification name are concatenated to determine the code corresponding to the target subject.
[0116] Situation 3: If there is only one target entry keyword and there are at least two target entry attribute keywords, then after determining the attribute priority of each target entry attribute classification, the codes corresponding to the names of each target entry attribute classification are concatenated according to the attribute priority to determine a code sequence (for convenience of description, denoted as the second code sequence). According to the preset code concatenation order, the code corresponding to the target entry name and the second code sequence are concatenated to determine the code corresponding to the target subject.
[0117] Situation 4: If there are at least two target entry keywords and there are at least two target entry attribute keywords, then after determining the entry priority of each target entry and the attribute priority of each target entry attribute classification, the codes corresponding to the names of each target entry are concatenated according to the entry priority to determine the first code sequence, and the codes corresponding to the names of each target entry attribute classification are concatenated according to the attribute priority to determine the second code sequence. According to the preset code concatenation order, the first code sequence and the second code sequence are concatenated to determine the code corresponding to the target subject.
[0118] Among them, the codes corresponding to each entry attribute classification name can be configured in the way of Cartesian product. Figure 2 It is a schematic diagram for configuring the codes corresponding to each entry attribute classification name provided by an embodiment of the present invention. As Figure 2As shown, for the n item attribute classifications included in any primary or secondary item, determine each detail item included in the item attribute classification, assign corresponding code values to the names of each detail item, and obtain the set of code values corresponding to the item attribute classification. For example, when financial assets are the primary or secondary item, the item attribute classifications included in the primary or secondary item include financial instrument classification, investment variety classification, listed and traded status classification, and accounting classification. For the names of the detail items included in the financial instrument classification, they are FVPL, FVOCI, and AC respectively, and the corresponding code values for the names of each detail item are 01, 02, and 03 respectively. Then, through the Cartesian product, process the code values corresponding to the names of the detail items included in each of the n item attribute classifications, so as to summarize the code values corresponding to the names of the detail items included in each of the n item attribute classifications under the primary or secondary item.
[0119] It can be determined that first, determine each item attribute classification that may be included under the primary or secondary item. Then, based on the actual experience values, the staff determines whether there are at least two item attribute classifications that are not coexistent. If it is determined that there are at least two item attribute classifications that are not coexistent, then do not determine the at least two item attribute classifications as the item attribute classifications included in the primary or secondary item.
[0120] Through the above method, when determining the code corresponding to the target item, since it is determined based on the codes corresponding to each target item name and the codes corresponding to each target item attribute classification name, the determined code corresponding to the target item can more intuitively reflect the items and item attribute classifications included in the target item, improve the accuracy of determining the code corresponding to the target item, and avoid the influence of the format of the valuation table on the code corresponding to the item. Moreover, if the staff adjusts the codes corresponding to each item name and / or the codes corresponding to each item attribute classification name, only one adjustment operation is required to achieve the use of the adjusted codes when summarizing valuation tables in different formats, reducing the workload of the staff and reducing the manpower and material resources consumed for adjusting the codes.
[0121] After determining the item name of the target item based on the above embodiment, the standard item name corresponding to the target content data can be determined according to the item name of the target item, and the standard item code sequence corresponding to the target content data can be determined according to the code corresponding to the target item.
[0122] In a possible implementation manner, the item name of the target item can be directly determined as the standard item name corresponding to the target content data, or a separator can be added to the item name of the target item, and the item name after adding the separator can be determined as the standard item name.
[0123] In a possible implementation, the corresponding code of the target subject can be directly determined as the standard subject code sequence, or a separator can be added to the code corresponding to the target subject, and the code after adding the separator is determined as the standard subject code sequence.
[0124] Embodiment 3:
[0125] In order to further improve the information in the summarized valuation table, based on the above embodiments, in the embodiments of the present invention, the determining the summarized valuation table according to the respective headers and the content data of the respective headers includes:
[0126] Match the target content data with preset place keywords to determine the place keywords that match the target content data;
[0127] Determine the target trading place corresponding to the matched place keywords;
[0128] Determine the summarized valuation table according to the respective headers, the content data of the respective headers, and the target trading place corresponding to the target content data.
[0129] In the actual application process, the subject name may contain information about the trading place, so as to uniquely determine the asset information contained in the subject name according to this trading place information. Therefore, in order to accurately determine the asset information contained in the subject name, in the embodiments of the present invention, place keywords corresponding to each trading place are preset according to the place keywords that may be contained in the content data with the subject name as the header in valuation tables in different formats, such as the Shanghai Stock Exchange, the Shenzhen Stock Exchange, etc. After obtaining the target content data based on the above embodiments, the target content data can be matched with the preset place keywords. Specifically, for each preset place keyword, the target content data is matched with the place keyword, that is, it is determined whether the target content data contains the place keyword. If it is determined that the target content data matches the place keyword, it means that the target content data contains the place keyword, and then according to the pre-configured correspondence between the trading place and the place keyword, the target trading place corresponding to the place keyword is determined. If it is determined that the target content data does not match the place keyword, it means that the target content data does not contain the place keyword, and then the next place keyword is obtained.
[0130] In a possible implementation, there may be a situation where the target content data does not match each place keyword, then the pre-configured default trading place can be determined as the target trading place corresponding to the target content data.
[0131] After determining the target trading venue corresponding to the target content data, a summarized valuation table can be determined based on each table header, the content data of each table header, and the target trading venue corresponding to the target content data.
[0132] Embodiment 4:
[0133] To ensure the accuracy of the data in the summarized valuation table, based on the above embodiments, in an embodiment of the present invention, the method further includes:
[0134] If the summarized valuation table contains an asymmetric total item, verify the data format of the first value in the asymmetric total item; where the asymmetric total item is a total item in the summarized valuation table that is not a symmetric total item, and the symmetric total item has a corresponding table header in the summarized valuation table.
[0135] In an actual application scenario, since the data formats of the values in valuation tables of different formats may be different. For example, some values have a percentage data format, and some values have a decimal data format. After aggregating values with different data formats, the data format of the obtained processed value may be incorrect. For example, the processed value is 0.01, but it is impossible to determine whether the data format of this value is 0.01% or simply 0.01. And generally, the data formats of the values under the same table header are the same. After aggregating the values under the same table header, the data format of the obtained processed value is generally correct. Therefore, to ensure the accuracy of the data in the summarized valuation table, in an embodiment of the present invention, after summarizing the valuation table to be summarized based on the above embodiments, each total item included in the summarized valuation table can be obtained. Then, for each total item, determine whether the total item is not a symmetric total item, that is, whether the total item does not have a corresponding table header in the summarized valuation table. For example, total items corresponding one by one to table headers such as quantity, unit cost, cost, market value, valuation increment, etc. If it is determined that the total item is not a symmetric total item, it means that the total item may be obtained by aggregating data with different data formats, then the data format of the value (for convenience of description, denoted as the first value) in the asymmetric total item can be verified. If it is determined that the total item is a symmetric total item, it means that the total item may be obtained by aggregating data with the same data format, and the data format of the value in the total item is generally correct, then there is no need to verify the data format of the first value in the total item.
[0136] In a possible implementation manner, the method for verifying the data format of the first value in the asymmetric total item includes the following:
[0137] Method 1: In an actual application scenario, if the data format of the first value is correct, then in the valuation tables summarized before the summarized valuation table, the data format of the value in the asymmetric total item (for the sake of convenience of explanation, denoted as the second value) will generally be the same as the data format of the first value. Therefore, after determining that a certain total item is an asymmetric total item, the valuation table summarized before the summarized valuation table can be obtained, and the second value in the asymmetric total item can be obtained from the obtained valuation table. Then, it is determined whether the data format of the second value is the same as the data format of the first value. If the data format of the second value is the same as the data format of the first value, it indicates that the data format of the first value may be correct, and thus there is no need to process the data format of the first value. If the data format of the second value is different from the data format of the first value, it indicates that the data format of the first value may be incorrect, and thus the data format of the first value can be modified according to the data format of the second value.
[0138] In a possible implementation manner, in order to further ensure the accuracy of the data in the summarized valuation table, in the embodiment of the present invention, if the data format of the second value is different from the data format of the first value, the electronic device can also generate a notification message and send it to the intelligent device of the staff to notify the staff to perform manual verification on the data format of the first value.
[0139] It should be noted that, in order to further ensure the accuracy of the data in the summarized valuation table, one or more valuation tables summarized before the summarized valuation table can be obtained, and the second value in the asymmetric total item can be obtained from the obtained one or more valuation tables. Then, for each second value, it is determined whether the data format of the first value is the same as the data format of the second value.
[0140] Method 2: The staff can set the preset data format requirements corresponding to the asymmetric total item. For example, the data format of the value in a certain asymmetric total item is a percentage, etc. Therefore, after determining that a certain total item is an asymmetric total item, the preset data format requirements corresponding to the asymmetric total item can be obtained. Then, it is determined whether the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item. If it is determined that the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item, it indicates that the data format of the first value may be correct, and thus there is no need to process the data format of the first value. If it is determined that the data format of the first value does not meet the preset data format requirements corresponding to the asymmetric total item, it indicates that the data format of the first value may be incorrect, and thus the data format of the first value can be modified according to the preset data format requirements corresponding to the asymmetric total item.
[0141] Through the above method, it is possible to avoid manual verification of the data formats of the numerical values of each total item in the summarized valuation table, reduce the manual work, lower the human and material resources consumed in verifying the data formats of the numerical values of each total item, improve the accuracy of the data in the summarized valuation table, and improve the efficiency of verifying the data formats of the numerical values of each total item.
[0142] Embodiment 5:
[0143] The following uses specific embodiments to illustrate the valuation table summarization method provided by the embodiments of the present invention. Figure 3 It is a schematic diagram of the specific valuation table summarization process provided by the embodiments of the present invention. The process includes:
[0144] S301: Match the data in the valuation table to be summarized with the preset header keyword terms to determine each header included in the data and the content data of each header in the data.
[0145] S302: Obtain the target content data with the header of "subject name" from the content data of each header.
[0146] S303: Match the target content data with the preset item keyword terms and the preset item attribute keyword terms respectively to determine the target item keyword terms that match the target content data and the target item attribute keyword terms that match the target content data.
[0147] S304: Determine the subject name of the target subject according to the target item name of the item corresponding to the target item keyword term and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword term.
[0148] Among them, the process of specifically determining the subject name of the target subject according to the target item name of the item corresponding to the target item keyword term and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword term has been introduced in the above embodiments. For specific reference, see the descriptions of Case 1 to Case 4 in Embodiment 2 above. Repeated parts will not be elaborated.
[0149] S305: Determine the code corresponding to the target subject according to the code corresponding to the target item name and the code corresponding to the target item attribute classification name.
[0150] Among them, the process of specifically determining the code corresponding to the target subject according to the code corresponding to the target item name and the code corresponding to the target item attribute classification name has been introduced in the above embodiments. For specific reference, see the descriptions of Case 1 to Case 4 in Embodiment 2 above. Repeated parts will not be elaborated.
[0151] S306: Determine the standard subject name corresponding to the target content data according to the subject name of the target subject, and determine the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject.
[0152] S307: Update the target content data according to the standard subject name, and update the content data corresponding to the target content data in the content data with the subject code as the header according to the standard subject code sequence.
[0153] S308: Match the target content data with the preset place keyword to determine the place keyword that matches the target content data.
[0154] S309: Determine the target trading place corresponding to the matched place keyword.
[0155] S310: Determine the summarized valuation table according to each header, the content data of each header, and the target trading place corresponding to the target content data.
[0156] S311: If the summarized valuation table contains an asymmetric total item, verify the data format of the first value in the asymmetric total item.
[0157] Among them, the asymmetric total item is the total item that is not a symmetric total item in the summarized valuation table, and the symmetric total item has a corresponding header in the summarized valuation table.
[0158] In a possible implementation manner, the specific ways to verify the data format of the first value in the asymmetric total item include the following two:
[0159] Obtain the valuation table summarized before this summarized valuation table; obtain the second value in the asymmetric total item from the obtained valuation table; determine whether the data format of the second value is the same as the data format of the first value; or
[0160] Determine whether the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item.
[0161] Since the header keyword terms, subject keyword terms, the corresponding relationship between subject keyword terms and subjects, and the corresponding relationship between subjects and codes are pre-configured, during the process of summarizing the valuation tables, by matching the data in the valuation tables to be summarized with the pre-set header keyword terms, each header included in the valuation tables to be summarized and the content data of each header in the valuation tables to be summarized can be determined. Then, the target content data with the subject name as the header is matched with the pre-set subject keyword terms to determine the target subject keyword terms that match the target content data. According to the subject name of the target subject corresponding to the target subject keyword terms, the standard subject name corresponding to the target content data can be determined, avoiding the problem that the subject names of the same subject are inconsistent in valuation tables with different formats, resulting in the inability to summarize valuation tables with different formats. And according to the corresponding relationship between subjects and codes, the code corresponding to the target subject can be obtained, so as to determine the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject, also avoiding the problem that the subject codes of the same subject are inconsistent in valuation tables with different formats, resulting in the inability to summarize valuation tables with different formats. Subsequently, the target content data is updated according to the standard subject name, and the content data corresponding to the target content data in the content data with the subject code as the header is updated according to the standard subject code sequence. Then, according to each header and the content data of each header, the summarized valuation table is determined. Through the above method, valuation tables with different formats can be accurately summarized without the need for staff to pre-set a set of templates or a set of rules for each format of valuation table, reducing the workload of the staff and avoiding the influence of the efficiency of the staff in configuring templates or rules on the efficiency of summarizing the valuation tables, improving the efficiency of summarizing valuation tables with different formats.
[0162] Embodiment 6:
[0163] An embodiment of the present invention provides a device for summarizing valuation tables. Figure 4 FIG. is a schematic structural diagram of a device for summarizing valuation tables provided by an embodiment of the present invention. The device includes:
[0164] A first processing unit 41, configured to match the data in the valuation tables to be summarized with the pre-set header keyword terms to determine each header included in the data and the content data of each header in the data;
[0165] A second processing unit 42, configured to match the target content data with the subject name as the header with the pre-set subject keyword terms to determine the target subject keyword terms that match the target content data;
[0166] A third processing unit 43, configured to determine a standard subject name corresponding to the target content data according to the subject name of the target subject corresponding to the target subject keyword, and determine a standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject;
[0167] An update unit 44, configured to update the target content data according to the standard subject name, and update the content data corresponding to the target content data in the content data with the subject code as the header according to the standard subject code sequence;
[0168] A summarization unit 45, configured to determine a summarized valuation table according to each header and the content data of each header.
[0169] Further, the third processing unit 43 is specifically configured to, if the subject keyword includes an item keyword and an item attribute keyword, match the target content data with the item keyword and the item attribute keyword respectively, and determine a target item keyword that matches the target content data, and a target item attribute keyword that matches the target content data; determine the subject name of the target subject according to the target item name of the item corresponding to the target item keyword and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword.
[0170] Further, the third processing unit 43 is specifically configured to, if there are at least two target item keywords and at least two target item attribute keywords, determine the item priority corresponding to each target item keyword according to a preset item priority, and determine the attribute priority of the item attribute classification corresponding to each target item attribute keyword according to the priority of the item attribute classification included in the preset target item; splice each target item name according to each item priority to obtain a first splicing sequence; and splice each target item attribute classification name according to each attribute priority to obtain a second splicing sequence; determine the subject name of the target subject according to the first splicing sequence and the second splicing sequence.
[0171] Further, the third processing unit 43 is specifically configured to splice the codes corresponding to each target item name according to each item priority to obtain a first code sequence; and splice the codes corresponding to each target item attribute classification name according to each attribute priority to obtain a second code sequence; determine the code corresponding to the target subject according to the first code sequence and the second code sequence.
[0172] Further, the device further includes: a fourth processing unit;
[0173] The fourth processing unit is configured to match the target content data with preset venue keyword terms to determine the venue keyword terms that match the target content data; and determine the target trading venue corresponding to the matched venue keyword terms.
[0174] The summarizing unit 45 is specifically configured to determine the summarized valuation table according to the respective headers, the content data of the respective headers, and the target trading venue corresponding to the target content data.
[0175] Further, the device further includes: a verification unit;
[0176] The verification unit is configured to, if the summarized valuation table includes an asymmetric total item, verify the data format of the first value in the asymmetric total item; wherein, the asymmetric total item is a total item in the summarized valuation table that is not a symmetric total item, and the symmetric total item has a corresponding header in the summarized valuation table.
[0177] Further, the verification unit is specifically configured to obtain the valuation table summarized before the summarized valuation table; obtain the second value in the asymmetric total item from the obtained valuation table; determine whether the data format of the second value is the same as the data format of the first value; or determine whether the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item.
[0178] Since the header keyword terms, subject keyword terms, the correspondence between subject keyword terms and subjects, and the correspondence between subjects and codes are pre-configured, during the process of summarizing the valuation tables, by matching the data in the valuation tables to be summarized with the pre-set header keyword terms, each header included in the valuation tables to be summarized and the content data of each header in the valuation tables to be summarized can be determined. Then, the target content data with the header being the subject name is matched with the pre-set subject keyword terms to determine the target subject keyword terms that match the target content data. According to the subject name of the target subject corresponding to the target subject keyword terms, the standard subject name corresponding to the target content data can be determined, avoiding the problem that the subject names of the same subject in valuation tables with different formats are inconsistent, which leads to the inability to summarize valuation tables with different formats. And according to the correspondence between subjects and codes, the code corresponding to the target subject can be obtained, so that according to the code corresponding to the target subject, the standard subject code sequence corresponding to the target content data can be determined, also avoiding the problem that the subject codes of the same subject in valuation tables with different formats are inconsistent, which leads to the inability to summarize valuation tables with different formats. Subsequently, the target content data is updated according to the standard subject name, and the content data corresponding to the target content data in the content data with the header being the subject code is updated according to the standard subject code sequence. Then, according to each header and the content data of each header, the summarized valuation table is determined. Through the above method, valuation tables with different formats can be accurately summarized without the need for staff to pre-set a set of templates or a set of rules for each format of valuation table, reducing the workload of the staff and avoiding the impact of the efficiency of staff in configuring templates or rules on the efficiency of summarizing valuation tables, improving the efficiency of summarizing valuation tables with different formats.
[0179] Embodiment 7:
[0180] Based on the above embodiments, an embodiment of the present invention further provides an electronic device, Figure 5 which is a schematic structural diagram of an electronic device provided by an embodiment of the present invention, as Figure 5 shown, including: a processor 51, a communication interface 52, a memory 53, and a communication bus 54, wherein the processor 51, the communication interface 52, and the memory 53 complete mutual communication through the communication bus 54;
[0181] A computer program is stored in the memory 53. When the program is executed by the processor 51, the processor 51 is caused to execute the following steps:
[0182] Match the data in the valuation tables to be summarized with the pre-set header keyword terms to determine each header included in the data and the content data of each header in the data;
[0183] Match the target content data with the header of the subject name against the preset subject keyword terms to determine the target subject keyword terms that match the target content data;
[0184] Determine the standard subject name corresponding to the target content data according to the subject name of the target subject corresponding to the target subject keyword terms, and determine the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject;
[0185] Update the target content data according to the standard subject name, and update the content data corresponding to the target content data in the content data with the header of the subject code according to the standard subject code sequence;
[0186] Determine the summarized valuation table according to each header and the content data of each header.
[0187] Further, the processor 51 is specifically configured to, if the subject keyword terms include item keyword terms and item attribute keyword terms, match the target content data with the item keyword terms and the item attribute keyword terms respectively to determine the target item keyword terms that match the target content data and the target item attribute keyword terms that match the target content data; determine the subject name of the target subject according to the target item name of the item corresponding to the target item keyword terms and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword terms.
[0188] Further, the processor 51 is specifically configured to, if there are at least two target item keyword terms and at least two target item attribute keyword terms, determine the item priority of each item corresponding to each target item keyword term according to the preset item priority, and determine the attribute priority of each item attribute classification corresponding to each target item attribute keyword term according to the priority of the item attribute classification included in the preset target item; splice each target item name according to each item priority to obtain a first splicing sequence; and splice each target item attribute classification name according to each attribute priority to obtain a second splicing sequence; determine the subject name of the target subject according to the first splicing sequence and the second splicing sequence.
[0189] Further, the processor 51 is specifically configured to splice the codes corresponding to each of the target entry names according to each entry priority to obtain a first code sequence; and splice the codes corresponding to each of the target entry attribute classification names according to each attribute priority to obtain a second code sequence; and determine the code corresponding to the target subject according to the first code sequence and the second code sequence.
[0190] Further, the processor 51 is further configured to match the target content data with preset venue keyword terms, determine the venue keyword terms matching the target content data; determine the target trading venue corresponding to the matching venue keyword terms; and determine the summarized valuation table according to the respective headers, the content data of the respective headers, and the target trading venue corresponding to the target content data.
[0191] Further, the processor 51 is further configured to, if the summarized valuation table includes an asymmetric total item, check the data format of the first value in the asymmetric total item; wherein, the asymmetric total item is a total item in the summarized valuation table that is not a symmetric total item, and the symmetric total item has a corresponding header in the summarized valuation table.
[0192] Further, the processor 51 is specifically configured to obtain the valuation table summarized before the summarized valuation table; obtain the second value in the asymmetric total item from the obtained valuation table; determine whether the data format of the second value is the same as the data format of the first value; or determine whether the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item.
[0193] Since the principle of the above electronic device for solving problems is similar to that of the valuation table summarization method, the implementation of the above electronic device can refer to the embodiments of the method, and the repeated parts will not be elaborated.
[0194] The communication bus mentioned in the above electronic device can be a Peripheral Component Interconnect (PCI) bus, an Extended Industry Standard Architecture (EISA) bus, or the like. This communication bus can be divided into an address bus, a data bus, a control bus, etc. For the sake of convenience of representation, only a thick line is used in the figure to represent it, but it does not mean that there is only one bus or one type of bus. The communication interface 52 is used for communication between the above electronic device and other devices. The memory can include a Random Access Memory (RAM), and can also include a Non-Volatile Memory (NVM), such as at least one disk memory. Optionally, the memory can also be at least one storage device located far from the aforementioned processor.
[0195] The above processor can be a general-purpose processor, including a central processing unit, a Network Processor (NP), etc.; it can also be a Digital Signal Processing (DSP), an application specific integrated circuit, a field programmable gate array, or other programmable logic devices, discrete gate or transistor logic devices, discrete hardware components, etc.
[0196] Since the header keyword terms, subject keyword terms, the correspondence between subject keyword terms and subjects, and the correspondence between subjects and codes are pre-configured, during the process of summarizing the valuation tables, by matching the data in the valuation tables to be summarized with the pre-set header keyword terms, it is possible to determine each header included in the valuation tables to be summarized and the content data of each header in the valuation tables to be summarized. Then, the target content data with the header of the subject name is matched with the pre-set subject keyword terms to determine the target subject keyword terms that match the target content data. According to the subject name of the target subject corresponding to the target subject keyword terms, the standard subject name corresponding to the target content data can be determined, avoiding the problem that the subject names of the same subject are inconsistent in valuation tables of different formats, resulting in the inability to summarize valuation tables of different formats. And according to the correspondence between subjects and codes, the code corresponding to the target subject can be obtained, so that according to the code corresponding to the target subject, the standard subject code sequence corresponding to the target content data can be determined, also avoiding the problem that the subject codes of the same subject are inconsistent in valuation tables of different formats, resulting in the inability to summarize valuation tables of different formats. Subsequently, the target content data is updated according to the standard subject name, and the content data corresponding to the target content data in the content data with the header of the subject code is updated according to the standard subject code sequence. Then, according to each header and the content data of each header, the summarized valuation table is determined. Through the above method, it is possible to accurately summarize valuation tables of different formats without the need for staff to pre-set a set of templates or a set of rules for each format of valuation table, reducing the workload of the staff and avoiding the impact of the efficiency of staff in configuring templates or rules on the efficiency of summarizing valuation tables, improving the efficiency of summarizing valuation tables of different formats.
[0197] Embodiment 8:
[0198] Based on the above embodiments, an embodiment of the present invention further provides a computer-readable storage medium, in which a computer program executable by a processor is stored. When the program runs on the processor, the processor is caused to execute the following steps:
[0199] Match the data in the valuation tables to be summarized with the pre-set header keyword terms to determine each header included in the data and the content data of each header in the data;
[0200] Match the target content data with the header of the subject name with the pre-set subject keyword terms to determine the target subject keyword terms that match the target content data;
[0201] Determine the standard subject name corresponding to the target content data according to the subject name of the target subject corresponding to the target subject keyword, and determine the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject;
[0202] Update the target content data according to the standard subject name, and update the content data corresponding to the target content data in the content data with the subject code as the header according to the standard subject code sequence;
[0203] Determine the summarized valuation table according to each header and the content data of each header.
[0204] Furthermore, the subject keyword includes an item keyword and an item attribute keyword. Obtaining the subject name of the target subject corresponding to the target subject keyword includes:
[0205] Match the target content data with the item keyword and the item attribute keyword respectively to determine the target item keyword that matches the target content data and the target item attribute keyword that matches the target content data;
[0206] Determine the subject name of the target subject according to the target item name of the item corresponding to the target item keyword and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword.
[0207] Furthermore, the determining the subject name of the target subject according to the target item name of the item corresponding to the target item keyword and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword includes:
[0208] If there are at least two target item keywords and at least two target item attribute keywords, determine the item priority of the item corresponding to each target item keyword according to the preset item priority, and determine the attribute priority of the item attribute classification corresponding to each target item attribute keyword according to the priority of the item attribute classification included in the preset target item;
[0209] Concatenate each target item name according to each item priority to obtain a first concatenated sequence; and
[0210] Concatenate each target item attribute classification name according to each attribute priority to obtain a second concatenated sequence;
[0211] Determine the subject name of the target subject according to the first splicing sequence and the second splicing sequence.
[0212] Further, obtaining the code corresponding to the target subject includes:
[0213] Splice the codes corresponding to each of the target entry names according to each entry priority to obtain a first code sequence; and
[0214] Splice the codes corresponding to the classification names of each target entry attribute according to each attribute priority to obtain a second code sequence;
[0215] Determine the code corresponding to the target subject according to the first code sequence and the second code sequence.
[0216] Further, the determining the summarized valuation table according to each of the table headers and the content data of each table header includes:
[0217] Match the target content data with preset place keyword terms to determine the place keyword terms that match the target content data;
[0218] Determine the target trading venue corresponding to the matched place keyword terms;
[0219] Determine the summarized valuation table according to each of the table headers, the content data of each table header, and the target trading venue corresponding to the target content data.
[0220] Further, the method further includes:
[0221] If the summarized valuation table contains an asymmetric total item, verify the data format of the first value in the asymmetric total item; wherein, the asymmetric total item is a total item in the summarized valuation table that is not a symmetric total item, and the symmetric total item has a corresponding table header in the summarized valuation table.
[0222] Further, the verifying the data format of the first value in the asymmetric total item includes:
[0223] Obtain the valuation table summarized before the summarized valuation table; obtain the second value in the asymmetric total item from the obtained valuation table; determine whether the data format of the second value is the same as the data format of the first value; or
[0224] Determine whether the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item.
[0225] Since the principle of the above computer-readable storage medium for solving problems is similar to that of the valuation table summarization method, the implementation of the above computer-readable storage medium can refer to the embodiments of the method, and the repeated parts will not be described again.
[0226] Since the header keyword, subject keyword, the correspondence between the subject keyword and the subject, and the correspondence between the subject and the code are pre-configured, during the valuation table summarization process, by matching the data in the valuation table to be summarized with the preset header keyword, each header included in the valuation table to be summarized and the content data of each header in the valuation table to be summarized can be determined. Then, the target content data with the header being the subject name is matched with the preset subject keyword to determine the target subject keyword that matches the target content data. According to the subject name of the target subject corresponding to the target subject keyword, the standard subject name corresponding to the target content data can be determined, avoiding the problem that the subject names of the same subject in different formats of valuation tables are inconsistent, resulting in the inability to summarize different formats of valuation tables. And according to the correspondence between the subject and the code, the code corresponding to the target subject can be obtained, so as to determine the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject, also avoiding the problem that the subject codes of the same subject in different formats of valuation tables are inconsistent, resulting in the inability to summarize different formats of valuation tables. Subsequently, the target content data is updated according to the standard subject name, and the content data corresponding to the target content data in the content data with the header being the subject code is updated according to the standard subject code sequence. Then, according to each header and the content data of each header, the summarized valuation table is determined. Through the above method, different formats of valuation tables can be accurately summarized without the need for staff to pre-set a set of templates or a set of rules for each format of valuation table in advance, reducing the workload of the staff and avoiding the impact of the efficiency of staff configuring templates or rules on the efficiency of summarizing valuation tables, improving the efficiency of summarizing different formats of valuation tables.
[0227] Those skilled in the art should understand that the embodiments of the present application can be provided as a method, a system, or a computer program product. Therefore, the present application can take the form of a complete hardware embodiment, a complete software embodiment, or an embodiment combining software and hardware aspects. Moreover, the present application can take the form of a computer program product implemented on one or more computer-usable storage media (including but not limited to disk storage, CD-ROM, optical storage, etc.) containing computer-usable program code.
[0228] This application is described with reference to the flowcharts and / or block diagrams of methods, apparatus (systems), and computer program products according to the application. It should be understood that each flow and / or block in the flowchart and / or block diagram, and the combination of flows and / or blocks in the flowchart and / or block diagram, can be implemented by computer program instructions. These computer program instructions can be provided to the processor of a general-purpose computer, a special-purpose computer, an embedded processor, or other programmable data processing devices to produce a machine, such that the instructions executed by the processor of the computer or other programmable data processing devices produce means for implementing the functions specified in a process Figure 1 one process or multiple processes and / or blocks Figure 1 or means for implementing the functions specified in multiple blocks.
[0229] These computer program instructions can also be stored in a computer-readable memory that can direct a computer or other programmable data processing device to work in a specific manner, such that the instructions stored in the computer-readable memory produce a manufactured article including instruction means that implement the functions specified in a process Figure 1 one process or multiple processes and / or blocks Figure 1 or means for implementing the functions specified in multiple blocks.
[0230] These computer program instructions can also be loaded onto a computer or other programmable data processing device, such that a series of operational steps are executed on the computer or other programmable device to produce a computer-implemented process, and thus the instructions executed on the computer or other programmable device provide steps for implementing the functions specified in a process Figure 1 one process or multiple processes and / or blocks Figure 1 or means for implementing the functions specified in multiple blocks.
[0231] Obviously, those skilled in the art can make various changes and modifications to this application without departing from the spirit and scope of this application. Thus, if these modifications and variations of this application fall within the scope of the claims of this application and their equivalent technologies, this application is also intended to include these changes and modifications.
Claims
1. A method for summarizing valuation tables, characterized in that, The method includes: Matching the data in the valuation table to be summarized with preset header keyword terms to determine each header included in the data and the content data of each header in the data; Matching the target content data with the header of the subject name with preset subject keyword terms to determine the target subject keyword terms that match the target content data; wherein, the subject keyword terms include item keyword terms and item attribute keyword terms; Matching the target content data with the item keyword terms and the item attribute keyword terms respectively to determine the target item keyword terms that match the target content data and the target item attribute keyword terms that match the target content data; determining the subject name of the target subject according to the target item name of the item corresponding to the target item keyword term and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword term; determining the standard subject name corresponding to the target content data according to the subject name of the target subject corresponding to the target subject keyword term; If there are at least two of the target item keyword terms and at least two of the target item attribute keyword terms, then according to each item priority, splicing the codes corresponding to the respective target item names to obtain a first code sequence; and according to each attribute priority, splicing the codes corresponding to the respective target item attribute classification names to obtain a second code sequence; determining the code corresponding to the target subject according to the first code sequence and the second code sequence; determining the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject; Updating the target content data according to the standard subject name, and updating the content data corresponding to the target content data in the content data with the header of the subject code according to the standard subject code sequence; Determining the summarized valuation table according to each header and the content data of each header.
2. The method according to claim 1, wherein The determining the subject name of the target subject according to the target item name of the item corresponding to the target item keyword term and the target item attribute classification name of the item attribute classification corresponding to the target item attribute keyword term includes: If there are at least two of the target item keyword terms and at least two of the target item attribute keyword terms, then determining the item priority of the item corresponding to each target item keyword term according to the preset item priority, and determining the attribute priority of the item attribute classification corresponding to each target item attribute keyword term according to the priority of the item attribute classification included in the preset target item; Splicing the respective target item names according to each item priority to obtain a first splicing sequence; and Splicing the respective target item attribute classification names according to each attribute priority to obtain a second splicing sequence; Determining the subject name of the target subject according to the first splicing sequence and the second splicing sequence.
3. The method according to claim 1, characterized in that, Determining the summarized valuation table according to the respective header tables and the content data of the respective header tables, including: Matching the target content data with preset place keyword terms to determine the place keyword terms matching the target content data; Determining the target trading place corresponding to the matching place keyword terms; Determining the summarized valuation table according to the respective header tables, the content data of the respective header tables, and the target trading place corresponding to the target content data.
4. The method according to claim 1, characterized in that, The method further includes: If the summarized valuation table contains an asymmetric total item, verifying the data format of the first value in the asymmetric total item; wherein, the asymmetric total item is a total item that is not a symmetric total item in the summarized valuation table, and the symmetric total item has a corresponding header table in the summarized valuation table.
5. The method according to claim 4, characterized in that, The verifying the data format of the first value in the asymmetric total item includes: Obtaining the valuation table summarized before the summarized valuation table; obtaining the second value in the asymmetric total item from the obtained valuation table; determining whether the data format of the second value is the same as the data format of the first value; or Determining whether the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item.
6. A valuation table summarizing device, characterized in that, The device includes: A first processing unit, configured to match the data in the valuation table to be summarized with preset header keyword terms to determine the respective header tables included in the data and the content data of the respective header tables in the data; A second processing unit, configured to match the target content data with a header of the subject name with preset subject keyword terms to determine the target subject keyword terms matching the target content data; wherein, the subject keyword terms include entry keyword terms and entry attribute keyword terms; A third processing unit, configured to match the target content data with the entry keyword terms and the entry attribute keyword terms respectively to determine the target entry keyword terms matching the target content data and the target entry attribute keyword terms matching the target content data; determining the subject name of the target subject according to the target entry name of the entry corresponding to the target entry keyword term and the target entry attribute classification name of the entry attribute classification corresponding to the target entry attribute keyword term; determining the standard subject name corresponding to the target content data according to the subject name of the target subject corresponding to the target subject keyword term; if there are at least two target entry keyword terms and at least two target entry attribute keyword terms, splicing the codes corresponding to the respective target entry names according to each entry priority to obtain a first code sequence; and splicing the codes corresponding to the respective target entry attribute classification names according to each attribute priority to obtain a second code sequence; determining the code corresponding to the target subject according to the first code sequence and the second code sequence; determining the standard subject code sequence corresponding to the target content data according to the code corresponding to the target subject. An update unit, configured to update the target content data according to the standard subject name, and update the content data corresponding to the target content data in the content data with the subject code as the header according to the standard subject code sequence; A summarization unit, configured to determine a summarized valuation table according to the respective headers and the content data of the respective headers; 7. The device according to claim 6, characterized in that, The third processing unit is specifically configured to, if there are at least two of the target entry keyword terms and at least two of the target entry attribute keyword terms, determine the entry priority of the entry corresponding to each of the target entry keyword terms according to a preset entry priority, and determine the attribute priority of the entry attribute classification corresponding to each of the target entry attribute keyword terms according to the priority of the entry attribute classification included in the preset target entry; splice each of the target entry names according to each of the entry priorities to obtain a first splicing sequence; and splice each of the target entry attribute classification names according to each of the attribute priorities to obtain a second splicing sequence; determine the subject name of the target subject according to the first splicing sequence and the second splicing sequence; 8. The device according to claim 6, characterized in that, The apparatus further includes: a fourth processing unit; The fourth processing unit is configured to match the target content data with a preset venue keyword term to determine a venue keyword term that matches the target content data; and determine a target trading venue corresponding to the matched venue keyword term; The summarization unit is specifically configured to determine the summarized valuation table according to the respective headers, the content data of the respective headers, and the target trading venue corresponding to the target content data; 9. The device according to claim 6, characterized in that, The apparatus further includes: a verification unit; The verification unit is configured to, if the summarized valuation table includes an asymmetric total item, verify the data format of the first value in the asymmetric total item; wherein, the asymmetric total item is a total item in the summarized valuation table that is not a symmetric total item, and the symmetric total item has a corresponding header in the summarized valuation table; 10. The device according to claim 9, characterized in that, The verification unit is specifically configured to obtain a valuation table summarized before the summarized valuation table; obtain a second value in the asymmetric total item from the obtained valuation table; determine whether the data format of the second value is the same as the data format of the first value; or determine whether the data format of the first value meets the preset data format requirements corresponding to the asymmetric total item; 11. An electronic device, characterized in that, The electronic device includes a processor, and the processor is configured to implement the steps of the valuation table summarization method according to any one of claims 1-5 when executing a computer program stored in a memory.
12. A computer-readable storage medium, characterized in that, It stores a computer program, and the computer program, when executed by a processor, implements the steps of the valuation table summarization method according to any one of claims 1-5.
Citation Information
Patent Citations
Method and device for improving subject code combination identifier generation efficiency
CN111046642A
Account asset supervision method and device
CN111062816A