Commodity personalized recommendation method and system based on user behavior data analysis
By constructing a user behavior sequence and interest stability model, and combining it with a dynamic product collaboration network, a personalized recommendation sequence is generated, which solves the problem of inaccurate recommendations in existing technologies and achieves more accurate and personalized product recommendations.
Patent Information
- Application Number
- CN202511068221.7
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-07-31
- Publication Date
- 2025-11-07
- Estimated Expiration
- 2045-07-31
AI Technical Summary
Existing product recommendation methods fail to comprehensively consider factors such as user behavior sequences, interest stability, and dynamic correlation of product attributes, resulting in inaccurate and untimely recommendation results that cannot meet the increasingly diverse and personalized needs of users.
A user behavior sequence and interest stability model is constructed, the behavior transfer correlation degree of adjacent interactive behavior units and the dynamic correlation degree of product attributes are calculated, a set of potential user behavior paths is generated based on the dynamic product collaborative network, and personalized recommendation sequences are generated by combining the attribute correlation feature distribution of the path and the user's current interactive behavior.
By comprehensively capturing the temporal sequence characteristics and interest changes of user interactions with products, the accuracy and personalization of recommendations are improved, thus enhancing the user shopping experience.
Smart Images

Figure CN120912294A_ABST
Abstract
Description
TECHNICAL FIELD
[0001] The present application relates to the technical field of artificial intelligence, in particular to a commodity personalized recommendation method and system based on user behavior data analysis. BACKGROUND
[0002] In the current booming e-commerce, commodity personalized recommendation system has become a key tool to improve user experience and platform sales performance. Traditional commodity recommendation methods are mainly divided into two categories. One is content-based recommendation, which generates recommendation results by analyzing the attribute characteristics of commodities and the attribute matching degree of user historical preference commodities. However, the above method ignores the time series information of user behavior and the dynamic change of interest, and is difficult to capture the transfer of user interest and the generation of new interest. The other is collaborative filtering-based recommendation, which uses the similarity between users or the similarity between commodities for recommendation. However, the collaborative filtering method usually calculates similarity statically without considering the stability of user interest and the change of commodity attribute correlation with time, resulting in inaccurate and timely recommendation results.
[0003] In addition, few methods in the prior art can comprehensively consider user behavior sequence, interest stability and commodity attribute dynamic correlation to construct a recommendation model, which cannot meet the increasingly diversified and personalized needs of users. SUMMARY
[0004] In view of the above-mentioned problems, in combination with the first aspect of the present application, the present application embodiment provides a commodity personalized recommendation method based on user behavior data analysis, which comprises: constructing a user behavior sequence and interest stability model, the user behavior sequence comprising user and commodity interaction behavior units arranged in time sequence, each interaction behavior unit being associated with interaction type and commodity attribute information, and the interest stability model being used to quantify the interest persistence characteristics of users to commodity attributes; performing association modeling processing on the user behavior sequence, and calculating the behavior transfer correlation and commodity attribute dynamic correlation of adjacent interaction behavior units in combination with the interest stability model, the commodity attribute dynamic correlation being adjusted according to the change of user interest stability; constructing a dynamic commodity collaborative network based on the commodity attribute dynamic correlation, the nodes of the dynamic commodity collaborative network being commodities, the edges being commodity attribute dynamic correlations, and the node importance parameters being updated according to real-time user interaction behavior; performing path optimization mining processing in the dynamic commodity collaborative network, and generating a user potential behavior path set in combination with the behavior transfer correlation and the node importance parameters, the potential behavior path set comprising multiple paths with different attribute correlation characteristics; The potential behavior path set is analyzed, and a personalized commodity recommendation sequence is generated based on the attribute association feature distribution of the path and the current interaction behavior of the user.
[0005] In another aspect, the embodiment of the present application also provides a personalized commodity recommendation system based on user behavior data analysis, comprising a processor, a machine readable storage medium, the machine readable storage medium is connected with the processor, the machine readable storage medium is used for storing programs, instructions or codes, and the processor is used for executing the programs, instructions or codes in the machine readable storage medium to realize the above-mentioned method.
[0006] Based on the above aspects, by constructing the user behavior sequence and the interest stability model, the time sequence characteristics of the user and the commodity interaction behavior and the continuous change of the user's interest in the commodity attributes can be comprehensively and accurately captured. When performing the association modeling processing on the user behavior sequence, the interest stability model is combined to calculate the behavior transition association degree and the commodity attribute dynamic association degree of the adjacent interaction behavior units, so that the commodity attribute association degree can be adjusted in real time according to the stability of the user's interest, and more closely matches the actual change of the user's interest. The node importance parameter of the dynamic commodity collaborative network constructed based on the commodity attribute dynamic association degree can be updated according to the real-time interaction behavior of the user, which ensures the dynamics and real-time performance of the network. The path optimization mining processing is performed in the dynamic commodity collaborative network, and the user potential behavior path set generated by combining the behavior transition association degree and the node importance parameter can more accurately predict the future behavior trend of the user. The final generated personalized commodity recommendation sequence comprehensively considers the attribute association feature distribution of the path and the current interaction behavior of the user, greatly improves the accuracy and personalization degree of the recommendation, and effectively improves the shopping experience of the user. BRIEF DESCRIPTION OF DRAWINGS
[0007] Figure 1 is the execution flow diagram of the personalized commodity recommendation method based on user behavior data analysis provided by the embodiment of the present application.
[0008] Figure 2 is the schematic diagram of the exemplary hardware and software components of the personalized commodity recommendation system based on user behavior data analysis provided by the embodiment of the present application. DETAILED DESCRIPTION
[0009] The present application will be specifically described below in conjunction with the drawings of the specification, Figure 1 is the flow diagram of the personalized commodity recommendation method based on user behavior data analysis provided by an embodiment of the present application, and the personalized commodity recommendation method based on user behavior data analysis will be described in detail below.
[0010] Step S110: Construct a user behavior sequence and interest stability model. The user behavior sequence includes user and product interaction behavior units arranged in chronological order. Each interaction behavior unit is associated with interaction type and product attribute information. The interest stability model is used to quantify the sustained characteristics of user interest in product attributes.
[0011] In this embodiment, constructing a user behavior sequence can clearly present the interaction history between users and products, while the interest stability model can quantify the persistence of users' interest in different product attributes, helping to more accurately grasp users' interest preferences. The specific implementation process will be described in detail below.
[0012] Step S111: Collect the user's original interaction records during the entire interaction period. The original interaction records include the time of interaction, the products involved in the interaction, the type of interaction, and the duration of the interaction. The interaction types include browsing interaction, favorites interaction, add-to-cart interaction, and purchase interaction.
[0013] In practical applications, data acquisition systems can be used to collect raw interaction records from multiple data sources throughout the entire user interaction cycle. These data sources can include e-commerce platform log systems, mobile application event tracking data, etc. The interaction timestamps in the raw interaction records clearly indicate the chronological order of user interactions with products; the products involved in the interactions identify the specific products the user is interested in; the interaction types reflect different user actions on the products; and the interaction duration reflects the time and effort the user invests in that interaction. For example, in an online shopping platform, the system records information related to each time a user opens a product details page (browsing interaction), clicks the favorite button (favorite interaction), adds the product to the shopping cart (add to cart interaction), and completes the purchase (purchase interaction), including the specific time, the names of the products involved, the type of interaction, and the time the user spends on each action.
[0014] Step S112: Divide the original interaction records into behavioral units, merge the continuous interaction records of the same product within a preset time window into one interactive behavioral unit. The interaction type of the interactive behavioral unit is the set of all interaction types contained in the merged records, and the interaction duration is the total interaction duration of the merged records.
[0015] In order to analyze the user's interaction behavior more effectively, it is necessary to divide the original interaction record into behavior units. The setting of the preset time window is a key factor, which determines which continuous interaction records will be combined. In actual operation, a suitable time window can be determined according to business needs and data analysis experience. For example, for some high-frequency interaction goods, the time window can be set shorter; for low-frequency interaction goods, the time window can be appropriately extended. When there are continuous interaction records of the same goods within the preset time window, these records are combined into an interaction behavior unit. The interaction type of this unit is the set of all interaction types in the combined records, which can more comprehensively reflect the user's operation behavior on the goods. The interaction duration is the total interaction duration of the combined records, which reflects the user's total input time on the goods. For example, if a user first performs a browsing interaction and then a collection interaction on a certain mobile phone within a short period of time, these two records will be combined into an interaction behavior unit, and the interaction type will be the set of browsing interaction and collection interaction, and the interaction duration will be the sum of the two interaction times.
[0016] Step S113: Extract the product attribute information of the product involved in each interaction behavior unit, the product attribute information including product category attribute, product function attribute, product scene attribute and product style attribute, wherein the product scene attribute is used to indicate the use scene applicable to the product, and the product style attribute is used to indicate the design style characteristics of the product.
[0017] Each interaction behavior unit is associated with a specific product, and extracting the attribute information of these products helps to deeply understand the characteristics of the products and the interest preferences of the users. The product category attribute can clearly indicate the category to which the product belongs, such as electronic products, clothing, food, etc. The product function attribute describes various functions possessed by the product, such as the camera function of a mobile phone, the processing capability of a computer, etc., which are important basis for users to select products. The product scene attribute indicates the use scene applicable to the product, such as sports scene, office scene, leisure scene, etc., which reflects the practicality of the product in different environments. The product style attribute embodies the design style characteristics of the product, such as minimalist style, retro style, fashion style, etc., which meets the needs of users for the appearance and aesthetic aspects of the product. For example, for a sports watch, its product category attribute is electronic products-watches, its product function attribute may include heart rate monitoring, motion trajectory recording, etc., its product scene attribute is applicable to sports scene, and its product style attribute may be minimalist fashion style.
[0018] Step S114: Sort the divided interaction behavior units in chronological order according to the interaction occurrence time, generate a preliminary user behavior sequence, perform missing interaction verification on the preliminary user behavior sequence, supplement the interaction behavior units corresponding to the key product attributes caused by the interruption of the interaction records, and obtain a complete user behavior sequence.
[0019] The sorted interaction behavior units can form a clear timeline of user interaction with goods, i.e., a preliminary user behavior sequence. The preliminary user behavior sequence can intuitively show which goods the user interacted with at different time points, as well as the type and duration of the interaction, etc. However, due to various reasons such as system failure, network problems, etc., the interaction record may be interrupted, resulting in missing interaction behavior units corresponding to key goods attributes. In order to ensure the integrity and accuracy of the user behavior sequence, the preliminary user behavior sequence needs to be checked for missing interactions. In actual operation, by analyzing the user's historical interaction behavior patterns, the relevance between goods attributes, and the continuity of time series, etc., it can be determined whether there is a missing interaction behavior unit. If missing is found, according to the relevant rules and algorithms, the corresponding interaction behavior unit is supplemented, so as to obtain a complete user behavior sequence. For example, if in the user's interaction record, the interaction record of a certain type of popular goods is suddenly interrupted within a certain time period, and according to the user's historical preferences and the popularity of the goods, it is speculated that the user may have browsed the goods during this period, at this time, the corresponding interaction behavior unit can be supplemented to the user behavior sequence.
[0020] Step S115: Based on the complete user behavior sequence, an interest stability model is constructed, the frequency of the same goods attribute information appearing in different time windows in the complete user behavior sequence is extracted, the fluctuation coefficient of the frequency change is calculated, and the reciprocal of the fluctuation coefficient is taken as the interest stability parameter of the user to the goods attribute. The higher the interest stability parameter is, the more stable the user's interest persistence feature for the goods attribute is.
[0021] The purpose of constructing the interest stability model is to quantify the user's interest persistence feature for the goods attribute, so as to better understand the user's interest preferences and behavior patterns. The specific implementation process is as follows: Step S1151: The time span of the complete user behavior sequence is evenly divided into multiple continuous time windows, each time window has the same duration, and there is no overlap between the time windows.
[0022] In order to analyze the change of user interest in the commodity attribute in detail, the time span of the complete user behavior sequence needs to be reasonably divided. In actual operation, the length of the time window can be determined according to business requirements and data characteristics. For example, if the short-term change of user interest is analyzed, the time window can be set to be short; if the long-term user interest trend is analyzed, the time window can be set to be long. The time span is evenly divided into a plurality of continuous time windows, and the length of each time window is the same, and there is no overlap between the time windows, so that the data of each time window has independence and comparability. For example, the complete behavior sequence of a user in a month is divided into a plurality of time windows in units of days, and each window represents the interaction data of one day.
[0023] Step S1152: For each commodity attribute information, the number of interaction behavior units appearing in each time window is counted, and the ratio of the number to the total number of interaction behavior units in the time window is taken as the appearance frequency of the commodity attribute information in the time window.
[0024] In each divided time window, for each commodity attribute information, the number of interaction behavior units appearing is counted. The number of interaction behavior units reflects the degree of attention of the user to the commodity attribute in the time window. Then, the number is compared with the total number of interaction behavior units in the time window, and the ratio is calculated, which is the appearance frequency of the commodity attribute information in the time window. For example, in a time window of a day, there is a certain number of interaction behavior units, and the interaction behavior units related to a commodity attribute (such as the camera function of a mobile phone) are a certain number. By calculating the ratio of the two numbers, the appearance frequency of the commodity attribute in the day can be obtained.
[0025] Step S1153: The appearance frequencies of all time windows are arranged in time sequence to form a frequency sequence, and the mean value of the frequency sequence is calculated, which is the arithmetic mean of all appearance frequencies.
[0026] The appearance frequencies of each commodity attribute information in each time window are arranged in time sequence to form a frequency sequence. The frequency sequence can directly show the change of the appearance frequency of the commodity attribute information at different time points. The mean value of the frequency sequence, which is the arithmetic mean of all appearance frequencies, can reflect the average appearance frequency of the commodity attribute information in the entire time span. Through the calculation of the mean value, the overall attention degree of the user to the commodity attribute can be preliminarily understood. For example, for a commodity attribute, the appearance frequencies in a plurality of time windows are different values, and the sum of the values divided by the number of time windows can obtain the mean value of the appearance frequency of the commodity attribute.
[0027] Step S1154: square the difference between each frequency in the frequency sequence and the mean, sum all the squared values, and divide by the number of time windows to obtain the variance of the frequency change, and take the square root of the variance as the volatility coefficient of the frequency change.
[0028] To further analyze the stability of user interest in the commodity attribute, the volatility of the frequency sequence needs to be calculated. First, the square of the difference between each frequency in the frequency sequence and the mean is calculated. This squared value can amplify the difference between the frequency and the mean, highlighting the degree of volatility. Then, sum all the squared values and divide by the number of time windows to obtain the variance of the frequency change. The variance reflects the degree of dispersion of the frequency sequence. The larger the variance, the greater the volatility of the frequency, and the less stable the user's interest in the commodity attribute. Finally, take the square root of the variance as the volatility coefficient of the frequency change, which more intuitively reflects the volatility amplitude of the frequency. For example, for a frequency sequence of a commodity attribute, the square of the difference between each frequency and the mean is calculated, the sum of these squared values is divided by the number of time windows to obtain the variance, and the square root of the variance is taken to obtain the volatility coefficient.
[0029] Step S1155: take the reciprocal of the volatility coefficient as the interest stability parameter of the user for the commodity attribute information, and set the interest stability parameter to a preset maximum value if the volatility coefficient is zero. The value range of the interest stability parameter is consistent with the value range of the commodity attribute dynamic association degree.
[0030] Taking the reciprocal of the volatility coefficient as the interest stability parameter of the user for the commodity attribute information is because the smaller the volatility coefficient, the more stable the user's interest in the commodity attribute, and the larger the reciprocal. If the volatility coefficient is zero, it means that the frequency sequence has no volatility, and the user's interest in the commodity attribute is very stable. At this time, the interest stability parameter is set to a preset maximum value. At the same time, in order to ensure the consistency and reasonableness of subsequent calculations, the value range of the interest stability parameter needs to be consistent with the value range of the commodity attribute dynamic association degree. For example, if the value range of the commodity attribute dynamic association degree is from a minimum value to a maximum value, the interest stability parameter should also be valued within the same range.
[0031] Step S120: perform association modeling processing on the user behavior sequence, and calculate the behavior transition association degree of adjacent interaction behavior units and the commodity attribute dynamic association degree in combination with the interest stability model. The commodity attribute dynamic association degree adjusts with the change of user interest stability.
[0032] After the user behavior sequence and the interest stability model are constructed, the user behavior sequence needs to be associated and modeled to calculate the behavior transition correlation and the commodity attribute dynamic correlation between adjacent interaction behavior units. The behavior transition correlation reflects the possibility and strength of the user transferring from one interaction behavior unit to the next interaction behavior unit, and the commodity attribute dynamic correlation reflects the correlation degree between the commodity attributes involved in adjacent interaction behavior units, and the correlation degree is adjusted with the change of the user interest stability. The specific implementation process is as follows: Step S121: identify the interaction type and the interaction duration of each interaction behavior unit in the user behavior sequence, assign a basic behavior weight to the interaction behavior unit according to the interaction type, and the basic behavior weight increases with the increase of the decision depth of the interaction type, wherein the basic behavior weight of the purchase interaction is greater than that of the add-to-cart interaction, the basic behavior weight of the add-to-cart interaction is greater than that of the collection interaction, and the basic behavior weight of the collection interaction is greater than that of the browsing interaction.
[0033] Firstly, the interaction type and the interaction duration of each interaction behavior unit in the user behavior sequence need to be identified. The interaction type includes browsing interaction, collection interaction, add-to-cart interaction and purchase interaction, and different interaction types reflect different decision depths of the user on the commodity. The browsing interaction is usually the initial understanding of the user on the commodity, and the decision depth is shallow; the collection interaction indicates that the user has a certain interest in the commodity, and the decision depth increases; the add-to-cart interaction shows that the user tends to purchase the commodity, and the decision depth further deepens; and the purchase interaction is the final decision behavior of the user, and the decision depth is the deepest. According to the interaction type, a basic behavior weight is assigned to the interaction behavior unit, and the weight increases with the increase of the decision depth of the interaction type. For example, a relatively small basic behavior weight is assigned to the browsing interaction, and a relatively large basic behavior weight is assigned to the purchase interaction.
[0034] Step S122: correct the basic behavior weight based on the interaction duration, calculate the ratio of the interaction duration to the average interaction duration of the interaction type, take the ratio as a duration correction factor, and multiply the basic behavior weight by the duration correction factor to obtain a corrected behavior weight value.
[0035] Interaction duration is also an important factor affecting the degree of user interest in the commodity. In order to more accurately reflect the user's interest, it is necessary to correct the basic behavior weight based on the interaction duration. The ratio of the interaction duration to the average interaction duration of the interaction type is calculated, and this ratio is the duration correction factor. If the interaction duration is greater than the average interaction duration of the interaction type, it means that the user has invested more time in the interaction, and the interest in the commodity may be more intense, so the duration correction factor will be greater than 1; on the contrary, if the interaction duration is less than the average interaction duration, the duration correction factor will be less than 1. Multiply the basic behavior weight by the duration correction factor to get the corrected behavior weight value. For example, for a certain interaction behavior unit, its basic behavior weight has been allocated according to the interaction type, and the duration correction factor is obtained by calculating the ratio of its interaction duration to the average interaction duration of the interaction type, and then multiplying the two to get the corrected behavior weight value.
[0036] Step S123: Extract the interaction occurrence time of the adjacent two interaction behavior units in the user behavior sequence, calculate the time interval parameter, construct a time decay function based on the time interval parameter, the output value of the time decay function decreases with the increase of the time interval parameter, and the corrected behavior weight value is multiplied by the output value of the time decay function to obtain the time decay behavior weight.
[0037] There may be a certain time interval between the interaction behaviors of the user at different time points, and this time interval will affect the degree of association between adjacent interaction behavior units. Extract the interaction occurrence time of the adjacent two interaction behavior units in the user behavior sequence, and calculate the time interval parameter between them. Based on the time interval parameter, a time decay function is constructed, and the output value of the function will decrease with the increase of the time interval parameter. This is because the longer the time interval, the more likely the user's interest and decision will change, and the degree of association between adjacent interaction behavior units will also weaken accordingly. Multiply the corrected behavior weight value by the output value of the time decay function to get the time decay behavior weight. For example, for two adjacent interaction behavior units, calculate the time interval between their interaction occurrence times, and get an output value through the time decay function according to the interval parameter, and then multiply the corrected behavior weight value by the output value to get the time decay behavior weight.
[0038] Step S124: Call the interest stability model to obtain the interest stability parameters corresponding to the respective commodity attribute information of the adjacent interaction behavior units, and take the product of the two interest stability parameters as the attribute stability factor.
[0039] The interest stability model has quantified the interest persistence characteristics of users on different commodity attributes. In calculating the dynamic correlation degree of commodity attributes, the interest stability model needs to be called to obtain the interest stability parameters of the commodity attribute information of each adjacent interaction behavior unit. These two parameters reflect the interest stability degree of the user on the commodity attributes involved in the adjacent interaction behavior units. Multiply these two interest stability parameters to obtain the attribute stability factor. The attribute stability factor reflects the interest stability correlation degree between the commodity attributes of adjacent interaction behavior units. For example, for two adjacent interaction behavior units, obtain the interest stability parameters corresponding to their commodity attribute information, multiply these two parameters to obtain the attribute stability factor.
[0040] Step S125: Extract the commodity attribute information of the adjacent interaction behavior units, calculate the coincidence degree of the commodity category attributes, the semantic matching degree of the commodity function attributes, the scene matching degree of the commodity scene attributes, and the style similarity of the commodity style attributes, and weight sum the category coincidence degree, the function semantic matching degree, the scene matching degree, and the style similarity after assigning a preset weight to each to obtain the basic attribute correlation degree.
[0041] Step S1251: Split the commodity category attributes of the adjacent interaction behavior units into multiple category labels respectively, count the number of the same category labels in the two commodity category attributes, and take the ratio of the number of the same category labels to the total number of category labels of the two commodity category attributes as the coincidence degree of the commodity category attributes.
[0042] In calculating the coincidence degree of the commodity category attributes, the commodity category attributes of the adjacent interaction behavior units are respectively split into multiple category labels. These labels can more specifically describe the category information of the commodities. Then, the number of the same category labels in the two commodity category attributes is counted, which reflects the overlapping part of the two commodities in the category. Compare the number of the same category labels with the total number of category labels of the two commodity category attributes to calculate the ratio, which is the coincidence degree of the commodity category attributes. For example, for two adjacent interaction behavior units, their commodity category attributes are respectively split into a plurality of category labels, the number of the same labels is counted, and the ratio to the total number of labels is calculated to obtain the coincidence degree of the commodity category attributes.
[0043] Step S1252: Perform text preprocessing on the commodity function attributes of the adjacent interaction behavior units, remove stop words and retain core function description words, input the core function description words into a pre-trained semantic encoder to generate function attribute vectors, and calculate the cosine similarity between the two function attribute vectors as the semantic matching degree of the commodity function attributes.
[0044] For the functional attributes of the commodities of the adjacent interaction behavior units, first, text preprocessing is performed. Stop words are some words that frequently appear in the text but do not help much in semantic understanding, such as “of”, “is”, “and” and the like. Removing these stop words can make the text more concise and highlight the core functional description words. The core functional description words that are retained are input into a pre-trained semantic encoder, and the encoder converts these words into a functional attribute vector. The functional attribute vector can more accurately represent the semantic information of the functional attributes of the commodities. The cosine similarity between the two functional attribute vectors is calculated, and the cosine similarity can measure the degree of similarity in direction of the two vectors. The closer the value is to 1, the more matched the semantics of the functional attributes of the two commodities are. The cosine similarity is taken as the semantic matching degree of the functional attributes of the commodities. For example, the functional attributes of the commodities of the adjacent two interaction behavior units are subjected to text preprocessing, and the core functional description words are obtained after the stop words are removed. The words are input into a semantic encoder to generate a functional attribute vector, the cosine similarity between the two vectors is calculated, and the semantic matching degree of the functional attributes of the commodities is obtained.
[0045] Step S1253: Extracting scene labels contained in the scene attributes of the commodities, each scene label corresponding to a preset scene weight, and calculating the sum of the scene weights of the same scene labels in the two scene attributes of the commodities. The ratio of the sum of the scene weights to the total sum of the scene weights of all scene labels of the two scene attributes of the commodities is taken as the scene matching degree of the scene attributes of the commodities.
[0046] Extracting scene labels contained in the scene attributes of the commodities, each scene label corresponding to a preset scene weight, and the scene weight reflecting the importance of the scene in the use of the commodity. The sum of the scene weights of the same scene labels in the two scene attributes of the commodities is calculated, which reflects the comprehensive importance of the two commodities in the same scene. Then, the sum of the scene weights is compared with the total sum of the scene weights of all scene labels of the two scene attributes of the commodities, and the ratio is calculated. The ratio is the scene matching degree of the scene attributes of the commodities. For example, for two adjacent interaction behavior units, the scene labels of the scene attributes of the commodities are extracted, the sum of the scene weights of the same scene labels is calculated, and the ratio of the sum to the total sum of the scene weights of all scene labels is calculated to obtain the scene matching degree of the scene attributes of the commodities.
[0047] Step S1254: Converting the style attributes of the commodities into style feature vectors, the dimensions of the style feature vectors corresponding to a preset set of style dimensions, and the value of each dimension being a feature value of the style dimension. The Euclidean distance between the two style feature vectors is calculated, and the reciprocal of the Euclidean distance is taken as the style similarity of the style attributes of the commodities. If the Euclidean distance is zero, the style similarity is set to a preset maximum value.
[0048] The commodity style attribute is converted into a style feature vector, and the dimension of the style feature vector corresponds to a preset style dimension set. The value of each dimension is a characteristic value of the style dimension, and the characteristic values can quantify the style characteristics of the commodity. The Euclidean distance between two style feature vectors is calculated, and the Euclidean distance can measure the distance between two vectors in space. The closer the distance, the more similar the styles of the two commodities. The reciprocal of the Euclidean distance is taken as the style similarity of the commodity style attribute, because the smaller the Euclidean distance, the higher the similarity, and taking the reciprocal can more intuitively reflect the above relationship. If the Euclidean distance is zero, it means that the styles of the two commodities are completely the same, and at this time the style similarity is set to a preset maximum value. For example, the commodity style attributes of two adjacent interaction behavior units are converted into style feature vectors, the Euclidean distance between them is calculated, the reciprocal is taken to obtain the style similarity, and if the Euclidean distance is zero, it is set to a preset maximum value.
[0049] Step S126: Multiply the basic attribute correlation degree by the attribute stability factor to obtain a commodity attribute dynamic correlation degree, and multiply the time decay behavior weight by the commodity attribute dynamic correlation degree to obtain a behavior transition correlation degree of adjacent interaction behavior units.
[0050] The basic attribute correlation degree is multiplied by the attribute stability factor to obtain the commodity attribute dynamic correlation degree. The commodity attribute dynamic correlation degree not only considers the basic correlation degree between the attributes of adjacent interaction behavior units, but also considers the interest stability of the user. The time decay behavior weight is multiplied by the commodity attribute dynamic correlation degree to obtain the behavior transition correlation degree of adjacent interaction behavior units. The behavior transition correlation degree comprehensively considers the behavior weight of the user, the time interval, and the correlation degree between the attributes of the commodities, and more comprehensively reflects the possibility and strength of the user transferring from one interaction behavior unit to the next interaction behavior unit. For example, the basic attribute correlation degree is multiplied by the attribute stability factor to obtain the commodity attribute dynamic correlation degree, and the time decay behavior weight is multiplied by the commodity attribute dynamic correlation degree to obtain the behavior transition correlation degree of adjacent interaction behavior units.
[0051] Step S130: A dynamic commodity collaborative network is constructed based on the commodity attribute dynamic correlation degree, wherein the nodes of the dynamic commodity collaborative network are commodities, the edges are commodity attribute dynamic correlation degrees, and the node importance parameters are updated according to real-time interaction behaviors of the user.
[0052] The dynamic commodity collaborative network is a graph structure for representing the correlation between commodities, which takes commodities as nodes and commodity attribute dynamic correlation degrees as edges, and can intuitively show the correlation degree between commodities. At the same time, the node importance parameters are updated according to the real-time interaction behaviors of the user to reflect the importance of the commodity in the current interest of the user. The specific construction process is as follows: Step S131: Collect all the goods involved in the interaction behavior units in the user behavior sequence, remove the repeated goods to form a goods set, and take each good in the goods set as an initial node of the dynamic goods collaborative network.
[0053] All the goods involved in the interaction behavior units in the user behavior sequence are collected. Since the user may interact with the same good multiple times, it is necessary to remove the repeated goods to form a unique goods set. Each good in the goods set is taken as an initial node of the dynamic goods collaborative network, and these nodes constitute the basic elements of the network. For example, in the user's behavior sequence, multiple mobile phones, computers and other goods are involved, and after removing the repeated goods, these different goods are taken as the initial nodes of the dynamic goods collaborative network.
[0054] Step S132: For any two goods in the goods set, if there is at least one adjacent interaction behavior unit in the user behavior sequence that involves the two goods respectively, the dynamic correlation degree of the corresponding goods attribute of the adjacent interaction behavior unit is extracted as the initial weight of the edge connecting the two goods nodes, and if there are multiple connection edges between the two goods, the average value of all initial weights is taken as the final edge weight.
[0055] For any two goods in the goods set, check whether there is at least one adjacent interaction behavior unit in the user behavior sequence that involves the two goods respectively. If there is the above adjacent interaction behavior unit, the dynamic correlation degree of the corresponding goods attribute of the adjacent interaction behavior unit is extracted as the initial weight of the edge connecting the two goods nodes. The initial weight reflects the correlation degree between the two goods. If there are multiple connection edges between the two goods (i.e. there are multiple adjacent interaction behavior units involving the two goods respectively), the average value of all initial weights is taken as the final edge weight to more accurately represent the comprehensive correlation degree between the two goods. For example, for goods A and goods B, if there are multiple adjacent interaction behavior units involving them in the user behavior sequence, the dynamic correlation degree of the corresponding goods attribute of each adjacent interaction behavior unit is extracted, and the average value of these correlation degrees is calculated as the final weight of the edge connecting goods A and goods B.
[0056] Step S133: Initialize the node importance parameter for each node in the dynamic goods collaborative network, and the initial node importance parameter is the ratio of the number of interaction behavior units of the good appearing in the user behavior sequence to the total number of interaction behavior units.
[0057] An initial node importance parameter of each node in the dynamic commodity collaborative network is initialized, which reflects the relative importance of the commodity in the user behavior sequence. The initial node importance parameter is calculated by the ratio of the number of interaction units of the commodity in the user behavior sequence to the total number of interaction units. For example, for a certain commodity, there are a certain number of interaction units of the commodity in the user behavior sequence, and there are a certain number of total interaction units. The ratio of the two numbers is calculated to obtain the initial node importance parameter of the commodity.
[0058] Step S134: Set a node importance update period, and collect real-time interaction behaviors of users in each update period. The real-time interaction behaviors are interaction units generated by users in the current update period.
[0059] The node importance update period is set, which can be determined according to business requirements and the frequency of data updates. In each update period, real-time interaction behaviors of users are collected, which are interaction units generated by users in the current update period. By continuously collecting real-time interaction behaviors, changes in user interest can be reflected in time, so as to update the importance parameters of nodes. For example, the update period is set to one day, and interaction units generated by users in this day are collected every day.
[0060] Step S135: Update the node importance parameter of the corresponding node according to the commodity involved in the real-time interaction behavior. If the interaction type of the real-time interaction behavior is purchase interaction, the node importance parameter increases by a preset increment value, and if it is browsing interaction, the node importance parameter decreases by a preset decrement value. The increment value and the decrement value corresponding to the add-to-cart interaction and the collection interaction are between the purchase interaction and the browsing interaction.
[0061] The node importance parameter of the corresponding node is updated according to the commodity involved in the real-time interaction behavior. Different interaction types have different effects on the node importance parameter. If the interaction type of the real-time interaction behavior is purchase interaction, it means that the user has a high interest and purchase intention for the commodity, and the node importance parameter will increase by a preset increment value. If it is browsing interaction, it may be the user's preliminary understanding, and the node importance parameter will decrease by a preset decrement value. The add-to-cart interaction and the collection interaction indicate that the user has a certain interest in the commodity, but it has not reached the purchase level, and the corresponding increment value and decrement value are between the purchase interaction and the browsing interaction. For example, when the user performs a purchase interaction, a certain commodity node is involved, and the importance parameter of the node will increase by a preset increment value. If it is browsing interaction, the node importance parameter will decrease by a preset decrement value.
[0062] Step S1351: Set corresponding weight adjustment coefficients for different interaction types, wherein the weight adjustment coefficient of the purchase interaction is a first coefficient, the weight adjustment coefficient of the add-to-cart interaction is a second coefficient, the weight adjustment coefficient of the collection interaction is a third coefficient, and the weight adjustment coefficient of the browsing interaction is a fourth coefficient, and the first coefficient > the second coefficient > the third coefficient > the fourth coefficient, and the fourth coefficient is a negative value.
[0063] The weight adjustment coefficients are set for different interaction types, and these coefficients are used to adjust the node importance parameter. The weight adjustment coefficient of the purchase interaction is the largest because it represents the final decision behavior of the user and the importance of the product is the largest; the weight adjustment coefficient of the browsing interaction is a negative value because it may be just a casual browsing of the user and the importance of the product is reduced to a certain extent; the weight adjustment coefficients of the add-to-cart interaction and the collection interaction are between the two. For example, a larger weight adjustment coefficient is set for the purchase interaction, and a smaller negative weight adjustment coefficient is set for the browsing interaction.
[0064] Step S1352: Obtain the interaction duration of the real-time interaction behavior, and calculate the ratio of the interaction duration to the average interaction duration of the interaction type as a duration influence factor, and the value range of the duration influence factor is a preset interval.
[0065] The interaction duration of the real-time interaction behavior is obtained, which is compared with the average interaction duration of the interaction type, and the ratio is calculated, which is the duration influence factor. The duration influence factor reflects the situation that the time invested by the user in the interaction behavior is relative to the average time of the interaction type. The value range of the duration influence factor is a preset interval to ensure its rationality and effectiveness. For example, if the interaction duration of a purchase interaction of the user is longer, the duration influence factor will be larger accordingly; if the interaction duration is shorter, the duration influence factor will be smaller.
[0066] Step S1353: Multiply the weight adjustment coefficient and the duration influence factor to obtain a node importance adjustment value, wherein the node importance adjustment value of the purchase interaction, the add-to-cart interaction and the collection interaction is a positive value, and the node importance adjustment value of the browsing interaction is a negative value.
[0067] The weight adjustment coefficient and the duration influence factor are multiplied to obtain the node importance adjustment value. For the purchase interaction, the add-to-cart interaction and the collection interaction, since the weight adjustment coefficient is a positive value and the duration influence factor is also a positive value, the node importance adjustment value is a positive value, which will increase the importance of the node; for the browsing interaction, the weight adjustment coefficient is a negative value, and the node importance adjustment value is a negative value, which will reduce the importance of the node. For example, for a purchase interaction, the weight adjustment coefficient and the duration influence factor are multiplied to obtain a positive node importance adjustment value, which is used to increase the importance of the corresponding node.
[0068] Step S1354: Extracting the real-time interaction behavior involves the current node importance parameter of the commodity in the dynamic commodity collaborative network, adding the current node importance parameter and the node importance adjustment value to obtain an updated node importance parameter.
[0069] Extracting the real-time interaction behavior involves the current node importance parameter of the commodity in the dynamic commodity collaborative network, adding the current node importance parameter and the node importance adjustment value to obtain an updated node importance parameter. This updating process can timely reflect the influence of the real-time interaction behavior of the user on the importance of the commodity. For example, for a certain commodity node, its current node importance parameter is known, and by calculating the node importance adjustment value, the two are added to obtain an updated node importance parameter.
[0070] Step S1355: Boundary constraint processing is performed on the updated node importance parameter. If the updated node importance parameter is greater than the preset upper limit value, it is set to the preset upper limit value, and if it is less than the preset lower limit value, it is set to the preset lower limit value, to ensure that the node importance parameter is within the preset effective range.
[0071] In order to ensure the rationality and effectiveness of the node importance parameter, boundary constraint processing needs to be performed on the updated node importance parameter. The preset upper limit value and the preset lower limit value specify the effective range of the node importance parameter. If the updated node importance parameter is greater than the preset upper limit value, it is set to the preset upper limit value; if it is less than the preset lower limit value, it is set to the preset lower limit value. For example, if the updated node importance parameter exceeds the preset upper limit value, it is adjusted to the preset upper limit value to ensure that the parameter is within a reasonable range.
[0072] Step S136: Periodic normalization processing is performed on the weights of all edges in the dynamic commodity collaborative network, so that the sum of the out-edge weights of each node is a fixed value. The normalization processing period is consistent with the node importance updating period.
[0073] Periodic normalization processing is performed on the weights of all edges in the dynamic commodity collaborative network, with the purpose of making the sum of the out-edge weights of each node a fixed value. This can ensure that the weights of the edges in the network are within a reasonable range, avoiding the situation of excessively large or small weights. The normalization processing period is consistent with the node importance updating period to ensure that the weights of the edges and the importance parameters of the nodes can be updated synchronously. For example, after each node importance updating period ends, normalization processing is performed on the weights of all edges in the network to keep the sum of the out-edge weights of each node fixed.
[0074] Step S140: performing path optimization mining processing in the dynamic commodity collaborative network, combining the behavior transition correlation degree and the node importance parameter to generate a user potential behavior path set, the potential behavior path set including multiple paths with different attribute correlation characteristics.
[0075] After the dynamic commodity collaborative network is constructed, path optimization mining processing needs to be performed therein to generate a user potential behavior path set. These potential behavior paths can reflect the possible behavior trajectory of the user under the current interest. The specific mining process is as follows: Step S141: determining the commodity involved in the last interaction behavior unit in the user behavior sequence as the starting node of path mining, and extracting the node importance parameter of the starting node in the dynamic commodity collaborative network as the path starting weight.
[0076] The commodity involved in the last interaction behavior unit in the user behavior sequence is determined as the starting node of path mining, because the commodity is the latest commodity concerned by the user, and is likely to be the focus of the current interest of the user. The node importance parameter of the starting node in the dynamic commodity collaborative network is extracted as the path starting weight, which reflects the importance of the starting node in the current interest of the user. For example, in the user's behavior sequence, the last interaction involves a latest smart watch, and the smart watch is taken as the starting node of path mining, and the node importance parameter thereof in the dynamic commodity collaborative network is extracted as the path starting weight.
[0077] Step S142: taking the starting node as the starting point, traversing all adjacent nodes directly connected to the starting node in the dynamic commodity collaborative network, taking the adjacent nodes as candidate next-hop nodes, and collecting the node importance parameter of each candidate next-hop node and the commodity attribute dynamic correlation degree of the connection edge.
[0078] Taking the starting node as the starting point, all adjacent nodes directly connected to the starting node in the dynamic commodity collaborative network are found out. These adjacent nodes are taken as candidate next-hop nodes, which are the commodities that the user may focus on next. The node importance parameter of each candidate next-hop node and the commodity attribute dynamic correlation degree of the connection edge are collected, which will be used for subsequent path selection. For example, for the starting node (smart watch), other commodity nodes (such as mobile phones, earphones, etc.) directly connected to it in the network are found out as candidate next-hop nodes, and their node importance parameters and commodity attribute dynamic correlation degrees of the connection edge are collected.
[0079] Step S143: for each candidate next-hop node, extracting the behavior transition correlation degree of the interaction behavior unit corresponding to the starting node in the user behavior sequence, multiplying the behavior transition correlation degree, the node importance parameter of the candidate next-hop node, and the commodity attribute dynamic correlation degree of the connection edge to obtain the path transition probability.
[0080] For each candidate next-hop node, the behavior transition correlation degree of the interaction behavior unit corresponding to the starting node in the user behavior sequence is extracted. The behavior transition correlation degree reflects the possibility of the user transitioning from the interaction behavior corresponding to the starting node to the interaction behavior corresponding to the candidate next-hop node. The behavior transition correlation degree, the node importance parameter of the candidate next-hop node, and the commodity attribute dynamic correlation degree of the connection edge are multiplied to obtain the path transition probability. The path transition probability comprehensively considers the possibility of behavior transition, the importance of the node, and the correlation degree between commodities, and can more accurately evaluate the possibility of the user selecting the path. For example, for a certain candidate next-hop node (mobile phone), the behavior transition correlation degree of the interaction behavior unit corresponding to the starting node (smart watch) is extracted, which is multiplied by the importance parameter of the mobile phone node and the commodity attribute dynamic correlation degree of the connection edge to obtain the path transition probability from the smart watch to the mobile phone.
[0081] Step S144: The candidate next-hop node with the highest path transition probability is selected as the second node of the path, and the node is taken as a new starting node. The steps of adjacent node traversal, path transition probability calculation, and node selection are repeatedly executed until the path length reaches a preset threshold or the adjacent node cannot be continuously traversed, and a preliminary potential behavior path is generated.
[0082] The candidate next-hop node with the highest path transition probability is selected as the second node of the path, which means that the user is most likely to select this path. The node is taken as a new starting node, and the steps of adjacent node traversal, path transition probability calculation, and node selection are repeatedly executed to continuously expand the path. The process will continue until the path length reaches a preset threshold or the adjacent node cannot be continuously traversed, at which time a preliminary potential behavior path is generated. For example, among the candidate next-hop nodes, the path transition probability of the mobile phone is the highest, the mobile phone is taken as the second node of the path, and then the mobile phone is taken as a new starting node to continue subsequent path mining until a termination condition is met, and a preliminary potential behavior path is generated.
[0083] Step S145: The preliminary potential behavior path is subjected to path pruning processing, and nodes with a node importance parameter lower than a preset node threshold and corresponding subsequent path segments are removed, and paths with all node importance parameters higher than the preset node threshold are retained as effective potential behavior paths.
[0084] The preliminary potential behavior path is pruned to remove nodes with low importance and corresponding path segments, so as to improve the quality and effectiveness of the path. The preset node threshold is a preset standard for judging whether the importance of a node is sufficient. If the node importance parameter of a node is lower than the preset node threshold, it means that the importance of the node in the current interest of the user is low, and the node and the corresponding subsequent path segment are removed. All paths with node importance parameters higher than the preset node threshold are retained as effective potential behavior paths. For example, in the preliminary potential behavior path, the node importance parameter of a node (such as a less popular accessory) is lower than the preset node threshold, and the node and the subsequent path segment are removed. The path that meets the conditions is retained as an effective potential behavior path.
[0085] Step S1451: Traverse each node in the preliminary potential behavior path, and extract the node importance parameter of the node in sequence according to the path order.
[0086] Each node in the preliminary potential behavior path is traversed, and the node importance parameter of the node is extracted in sequence according to the path order. This process can systematically check the importance of each node in the path. For example, starting from the starting node of the preliminary potential behavior path, the node importance parameter of each node is extracted in sequence.
[0087] Step S1452: Compare the node importance parameter of each node with the preset node threshold, and mark the first node whose node importance parameter is lower than the preset node threshold as a pruning node.
[0088] The node importance parameter of each node is compared with the preset node threshold, and the first node whose node importance parameter is lower than the preset node threshold is found and marked as a pruning node. The pruning node is the first node with insufficient importance in the path, and needs to be pruned. For example, during the traversal process, a node whose node importance parameter is lower than the preset node threshold is found and marked as a pruning node.
[0089] Step S1453: If there is no pruning node in the preliminary potential behavior path, the path is directly used as an effective potential behavior path.
[0090] If there is no pruning node in the preliminary potential behavior path, it means that all node importance parameters in the path are higher than the preset node threshold, and the path meets the requirements and can be directly used as an effective potential behavior path. For example, after comparison, all node importance parameters in the preliminary potential behavior path are higher than the preset node threshold, so the path is an effective potential behavior path.
[0091] Step S1454: If there is a pruning node, remove the pruning node and all path segments after the node, and keep the path segments before the pruning node as the candidate pruning path.
[0092] If there is a pruning node, remove the pruning node and all path segments after the node, and keep the path segments before the pruning node as the candidate pruning path. This operation can remove unimportant parts of the path and improve the quality of the path. For example, when a pruning node is found, the node and subsequent path segments are deleted, and the previous path segments are kept as the candidate pruning path.
[0093] Step S1455: Calculate the path length of the candidate pruning path, and if the path length is less than the preset minimum path length, discard the candidate pruning path, otherwise keep the candidate pruning path as an effective potential behavior path.
[0094] Calculate the path length of the candidate pruning path and compare it with the preset minimum path length. If the path length is less than the preset minimum path length, it means that the path is too short and may not provide enough information, so it is discarded; otherwise, the candidate pruning path is kept as an effective potential behavior path. For example, calculate the length of the candidate pruning path, and if the length is less than the preset minimum path length, discard the path; if the length meets the requirements, keep it as an effective potential behavior path.
[0095] Step S146: Repeat the above steps to generate multiple effective potential behavior paths, so that different potential behavior paths contain different combinations of associated features of product attributes, forming a set of user potential behavior paths.
[0096] Repeat the above path mining and pruning steps to generate multiple effective potential behavior paths. These paths contain different combinations of associated features of product attributes and can more comprehensively reflect the user's potential behavior trajectory. Collect these effective potential behavior paths to form a set of user potential behavior paths. For example, through multiple iterations of the path mining and pruning process, multiple different effective potential behavior paths are generated, which cover different combinations of products and associated features of attributes, and together form a set of user potential behavior paths.
[0097] Step S150: Analyze the set of potential behavior paths, and generate a personalized product recommendation sequence based on the distribution of the path's attribute association features and the user's current interaction behavior.
[0098] Analyze the generated set of potential behavior paths, combine the distribution of the path's attribute association features and the user's current interaction behavior, and generate a product recommendation sequence that meets the user's personalized needs. The specific analysis and generation process is as follows: For example, step S151: extract the product attribute association features of each path in the set of potential behavior paths, which include the sequence distribution of product category attributes, the combination features of product function attributes, the coverage range of product scene attributes, and the change trend of product style attributes.
[0099] The product attribute association features of each path in the set of potential behavior paths are extracted, which can reflect the characteristics and association of products in the path. The sequence distribution of product category attributes shows the distribution of product categories in the path in terms of order, reflecting the user's interest shift between different categories of products. The combination features of product function attributes reflect the matching and combination of product functions in the path, meeting the user's diverse needs in terms of function. The coverage range of product scene attributes represents the breadth and depth of the application scenarios of products in the path, which can adapt to the user's use needs in different scenarios. The change trend of product style attributes reflects the evolution of product styles in the path, embodying the user's preference for different styles. For example, in a certain potential behavior path, the sequence distribution of product category attributes may be from mobile phones to earphones to smart watches, the combination features of product function attributes may be a combination of functions such as taking pictures, playing music, and monitoring sports, the coverage range of product scene attributes may include daily use, sports scenarios, and office scenarios, and the change trend of product style attributes may be a shift from minimalist style to fashion style.
[0100] Step S152: Calculate the attribute association feature similarity between any two paths, which is determined by comparing the dynamic time warping distance of category sequence distribution, the Jaccard similarity coefficient of function combination features, the intersection-over-union ratio of scene coverage range, and the Pearson correlation coefficient of style change trend.
[0101] The attribute association feature similarity between any two paths needs to consider multiple aspects of features. For the category sequence distribution, dynamic time warping distance is used for comparison, which can handle different sequence lengths and more accurately measure the similarity between category sequences. For the function combination features, the Jaccard similarity coefficient is used for calculation, which can reflect the similarity between two sets. For the scene coverage range, the intersection-over-union ratio is used to measure, which can intuitively show the overlap between two scene coverage ranges. For the style change trend, the Pearson correlation coefficient is used to determine, which can measure the linear correlation between two variables. By integrating these indicators, the attribute association feature similarity between any two paths can be obtained. For example, for two potential behavior paths, the dynamic time warping distance of their category sequence distribution, the Jaccard similarity coefficient of their function combination features, the intersection-over-union ratio of their scene coverage range, and the Pearson correlation coefficient of their style change trend are calculated, and the attribute association feature similarity between them is obtained by integrating these indicators.
[0102] Step S153: Perform clustering processing on the set of potential behavior paths based on the attribute association feature similarity, and divide the paths with a similarity higher than a preset clustering threshold into the same path cluster. Each path cluster contains paths with similar attribute association features.
[0103] Perform clustering processing on the set of potential behavior paths based on the attribute association feature similarity, and divide the paths with a similarity higher than a preset clustering threshold into the same path cluster. The preset clustering threshold is a pre-set standard for determining whether the similarity of two paths is high enough. The paths in the same path cluster have similar attribute association features, which represent the similar interests and behavior patterns of users in some aspects. For example, divide the paths with a similarity higher than a preset clustering threshold into a path cluster, and the paths in the path cluster may all have similar commodity category sequence distribution, function combination feature, scene coverage range, and style change trend.
[0104] Step S154: Select the path with the highest path comprehensive score from each path cluster as the representative path, and the path comprehensive score is obtained by multiplying the reciprocal of the path length by the cumulative value of the commodity attribute dynamic association degree in the path.
[0105] Select the path with the highest path comprehensive score from each path cluster as the representative path, and the calculation method of the path comprehensive score is the multiplication of the reciprocal of the path length and the cumulative value of the commodity attribute dynamic association degree in the path. The reciprocal of the path length reflects the simplicity of the path, and the shorter the path, the larger the reciprocal, indicating that the path is more efficient. The cumulative value of the commodity attribute dynamic association degree in the path reflects the degree of association between commodities in the path, and the larger the cumulative value, the closer the association between commodities. By multiplying the two, the simplicity of the path and the degree of association between commodities can be considered comprehensively to select the optimal representative path. For example, for each path cluster, calculate the comprehensive score of each path in the path cluster, and select the path with the highest score as the representative path.
[0106] Step S155: Collect all representative paths, extract the commodity nodes in the representative paths, and arrange them in the order of the paths to form an initial recommendation sequence.
[0107] Collect all representative paths of the path clusters, extract the commodity nodes from these representative paths, and arrange them in the order of the paths to form an initial recommendation sequence. The commodities in the initial recommendation sequence are selected based on the potential behavior paths and attribute association features of the user, and have certain relevance and rationality. For example, arrange the commodity nodes in all representative paths in the order of the paths to form an initial commodity recommendation sequence.
[0108] Step S156: Collect the current interaction behavior of the user, the current interaction behavior being the interaction behavior unit generated by the user within a preset time before generating the recommendation sequence, and extract the product attribute information of the product involved in the current interaction behavior.
[0109] The current interaction behavior of the user is collected, which is the interaction behavior unit generated by the user within a preset time before generating the recommendation sequence. The product attribute information of the product involved in the current interaction behavior is extracted, which can reflect the current interest and demand of the user. For example, within a period of time before generating the recommendation sequence, the user interacts with a certain mobile phone, and the product attribute information of the mobile phone is extracted, including category attribute, function attribute, scene attribute and style attribute, etc.
[0110] Step S157: Filtering the initial recommendation sequence according to the product attribute information of the current interaction behavior, removing the product nodes whose attribute correlation with the product attribute information of the current interaction behavior is lower than a preset filtering threshold, and then sorting the filtered product nodes according to the order in the representative path and the path comprehensive score to generate the final personalized product recommendation sequence.
[0111] The initial recommendation sequence is filtered according to the product attribute information of the current interaction behavior. The preset filtering threshold is a pre-set standard for judging whether the attribute correlation between products is high enough. If the attribute correlation between a certain product node in the initial recommendation sequence and the product attribute information of the current interaction behavior is lower than the preset filtering threshold, it means that the product is not closely related to the current interest of the user, and it is removed. Then, the filtered product nodes are sorted according to the order in the representative path and the path comprehensive score to generate the final personalized product recommendation sequence. The personalized product recommendation sequence can better meet the personalized needs of the user and provide product recommendations that better meet the current interest of the user. For example, for the product nodes in the initial recommendation sequence, the attribute correlation between them and the product attribute information of the current interaction behavior is calculated, the nodes with correlation lower than the preset filtering threshold are removed, and the remaining nodes are sorted to generate the final personalized product recommendation sequence.
[0112] Figure 2 A schematic diagram of exemplary hardware and software components of the product personalized recommendation system 100 based on user behavior data analysis that can implement the idea of the present application is shown. For example, the processor 120 can be used in the product personalized recommendation system 100 based on user behavior data analysis and used to perform the functions in the present application.
[0113] The user behavior data analysis based commodity personalized recommendation system 100 can be a general server or a special purpose server, both of which can be used to implement the user behavior data analysis based commodity personalized recommendation method of the present application. The present application only shows one server, but for the sake of convenience, the functions described in the present application can be implemented in a distributed manner on multiple similar platforms to balance the processing load.
[0114] For example, the user behavior data analysis based commodity personalized recommendation system 100 can include a network port 110 connected to a network, one or more processors 120 for executing program instructions, a communication bus 130, and different forms of storage media 140, such as a disk, a ROM, or a RAM, or any combination thereof. The user behavior data analysis based commodity personalized recommendation system 100 can also include program instructions stored in a ROM, a RAM, or other types of non-transitory storage media, or any combination thereof, for example. The method of the present application can be implemented according to these program instructions. The user behavior data analysis based commodity personalized recommendation system 100 also includes an input / output (I / O) interface 150 between the computer and other input / output devices.
[0115] For the sake of illustration, only one processor is described in the user behavior data analysis based commodity personalized recommendation system 100. However, it should be noted that the user behavior data analysis based commodity personalized recommendation system 100 in the present application can also include multiple processors, so the steps performed by one processor described in the present application can also be jointly performed or separately performed by multiple processors. For example, if the processor of the user behavior data analysis based commodity personalized recommendation system 100 performs steps A and B, it should be understood that steps A and B can also be jointly performed by two different processors or separately performed in one processor. For example, a first processor performs step A and a second processor performs step B, or the first processor and the second processor jointly perform steps A and B.
[0116] In addition, the present application also provides a readable storage medium, in which computer executable instructions are pre-set, and when the processor executes the computer executable instructions, the user behavior data analysis based commodity personalized recommendation method described above is implemented.
[0117] It should be noted that, in order to simplify the description of the present application and to help understand one or more embodiments of the present application, in the foregoing description of the embodiments of the present application, various features are sometimes combined into one embodiment, figure or description thereof.
Claims
1. A method for personalized recommendation of goods based on user behavior data analysis, characterized in that, The method comprises: constructing a user behavior sequence and interest stability model, the user behavior sequence comprising user interaction behavior units with goods arranged in chronological order, each interaction behavior unit being associated with an interaction type and a good attribute information, and the interest stability model being used for quantifying a user's interest persistence characteristic for a good attribute; performing associated modeling processing on the user behavior sequence, calculating a behavior transition association degree and a good attribute dynamic association degree of adjacent interaction behavior units in combination with the interest stability model, and the good attribute dynamic association degree being adjusted according to a user interest stability change; constructing a dynamic good collaborative network based on the good attribute dynamic association degree, the nodes of the dynamic good collaborative network being goods, the edges being good attribute dynamic association degrees, and a node importance parameter being updated according to a user real-time interaction behavior; performing path optimization mining processing in the dynamic good collaborative network, generating a user potential behavior path set in combination with the behavior transition association degree and the node importance parameter, and the potential behavior path set comprising multiple paths with different attribute association characteristics; analyzing the potential behavior path set, and generating a good personalized recommendation sequence based on a path attribute association characteristic distribution and a user current interaction behavior. 2.The method of claim 1, wherein, The method comprises: collecting original interaction records of a user in a full interaction period, the original interaction records comprising an interaction occurrence time, an interaction involved good, an interaction type and an interaction duration, and the interaction type comprising a browsing interaction, a collection interaction, an addition interaction and a purchase interaction; dividing the original interaction records into behavior units, and merging continuous interaction records of a same good in a preset time window into one interaction behavior unit, the interaction type of the interaction behavior unit being a set of all interaction types contained in the merged records, and the interaction duration being a total interaction duration of the merged records; extracting good attribute information of a good involved in each interaction behavior unit, the good attribute information comprising a good category attribute, a good function attribute, a good scene attribute and a good style attribute, wherein the good scene attribute is used for indicating a use scene applicable to the good, and the good style attribute is used for indicating a design style characteristic of the good; sorting the divided interaction behavior units in a chronological order of the interaction occurrence time, generating a preliminary user behavior sequence, and performing missing interaction verification on the preliminary user behavior sequence, supplementing an interaction behavior unit corresponding to a key good attribute caused by an interaction record interruption, and obtaining a complete user behavior sequence; constructing an interest stability model based on the complete user behavior sequence, extracting an appearance frequency of same good attribute information in different time windows in the complete user behavior sequence, calculating a fluctuation coefficient of the frequency change, and taking a reciprocal of the fluctuation coefficient as an interest stability parameter of the user for the good attribute, and the higher the interest stability parameter is, the more stable a user's interest persistence characteristic for the good attribute is. 3.The method of claim 2, wherein, The method comprises: collecting original interaction records of a user in a full interaction period, the original interaction records comprising an interaction occurrence time, an interaction involved good, an interaction type and an interaction duration, and the interaction type comprising a browsing interaction, a collection interaction, an addition interaction and a purchase interaction; dividing the original interaction records into behavior units, and merging continuous interaction records of a same good in a preset time window into one interaction behavior unit, the interaction type of the interaction behavior unit being a set of all interaction types contained in the merged records, and the interaction duration being a total interaction duration of the merged records; extracting good attribute information of a good involved in each interaction behavior unit, the good attribute information comprising a good category attribute, a good function attribute, a good scene attribute and a good style attribute, wherein the good scene attribute is used for indicating a use scene applicable to the good, and the good style attribute is used for indicating a design style characteristic of the good; sorting the divided interaction behavior units in a chronological order of the interaction occurrence time, generating a preliminary user behavior sequence, and performing missing interaction verification on the preliminary user behavior sequence, supplementing an interaction behavior unit corresponding to a key good attribute caused by an interaction record interruption, and obtaining a complete user behavior sequence; constructing an interest stability model based on the complete user behavior sequence, extracting an appearance frequency of same good attribute information in different time windows in the complete user behavior sequence, calculating a fluctuation coefficient of the frequency change, and taking a reciprocal of the fluctuation coefficient as an interest stability parameter of the user for the good attribute, and the higher the interest stability parameter is, the more stable a user's interest persistence characteristic for the good attribute is. The time span of the complete user behavior sequence is evenly divided into multiple continuous time windows, each time window has the same length, and there is no overlap between the time windows; For each item attribute information, the number of interaction behavior units appearing in each time window is counted, and the ratio of the number to the total number of interaction behavior units in the time window is taken as the appearance frequency of the item attribute information in the time window; The appearance frequencies of all time windows are arranged in time sequence to form a frequency sequence, and the mean value of the frequency sequence is calculated, which is the arithmetic mean of all appearance frequencies; The square of the difference between each appearance frequency in the frequency sequence and the mean value is calculated, and the sum of all square values is divided by the number of time windows to obtain the variance value of the frequency change, and the square root of the variance value is taken as the fluctuation coefficient of the frequency change; The reciprocal of the fluctuation coefficient is taken as the interest stability parameter of the user for the item attribute information, and if the fluctuation coefficient is zero, the interest stability parameter is set to a preset maximum value, and the value range of the interest stability parameter is consistent with the value range of the item attribute dynamic association degree. 4.The method of claim 1, wherein, The association modeling processing is performed on the user behavior sequence, and the behavior transition association degree and the item attribute dynamic association degree of adjacent interaction behavior units are calculated in combination with the interest stability model, including: The interaction type and interaction duration of each interaction behavior unit in the user behavior sequence are identified, and a basic behavior weight is assigned to the interaction behavior unit according to the interaction type, and the basic behavior weight increases with the increase of the decision depth of the interaction type, wherein the basic behavior weight of the purchase interaction is greater than that of the add-to-cart interaction, the basic behavior weight of the add-to-cart interaction is greater than that of the collection interaction, and the basic behavior weight of the collection interaction is greater than that of the browsing interaction; The basic behavior weight is corrected based on the interaction duration, the ratio of the interaction duration to the average interaction duration of the interaction type is taken as a duration correction factor, and the basic behavior weight and the duration correction factor are multiplied to obtain a corrected behavior weight value; The interaction occurrence time of adjacent two interaction behavior units in the user behavior sequence is extracted, a time interval parameter is calculated, and a time decay function is constructed based on the time interval parameter, wherein the output value of the time decay function decreases with the increase of the time interval parameter, and the corrected behavior weight value and the output value of the time decay function are multiplied to obtain a time decay behavior weight; The interest stability parameters corresponding to the item attribute information of each adjacent interaction behavior unit are obtained by calling the interest stability model, and the product of the two interest stability parameters is taken as an attribute stability factor; The item attribute information of adjacent interaction behavior units is extracted, the coincidence degree of the item category attribute, the semantic matching degree of the item function attribute, the scene matching degree of the item scene attribute, and the style similarity of the item style attribute are calculated, and the category coincidence degree, the function semantic matching degree, the scene matching degree, and the style similarity are weighted and summed after being assigned with preset weights to obtain a basic attribute association degree; The basic attribute association degree and the attribute stability factor are multiplied to obtain the item attribute dynamic association degree, and the time decay behavior weight and the item attribute dynamic association degree are multiplied to obtain the behavior transition association degree of adjacent interaction behavior units. 5.The method of claim 4, wherein, The method comprises the following steps: The commodity category attributes of the adjacent interaction behavior units are respectively split into multiple category labels, the number of same category labels in the two commodity category attributes is counted, and the ratio of the number of same category labels to the total number of category labels of the two commodity category attributes is taken as the coincidence degree of the commodity category attributes; The commodity function attributes of the adjacent interaction behavior units are subjected to text preprocessing, the core function description words are reserved after the stop words are removed, the core function description words are input into a pre-trained semantic encoder to generate function attribute vectors, and the cosine similarity between the two function attribute vectors is taken as the semantic matching degree of the commodity function attributes; The scene labels contained in the commodity scene attributes are extracted, each scene label corresponds to a preset scene weight, the sum of the scene weights of the same scene labels in the two commodity scene attributes is counted, and the ratio of the sum of the scene weights to the total sum of the scene weights of all scene labels of the two commodity scene attributes is taken as the scene fit degree of the commodity scene attributes; The commodity style attributes are converted into style feature vectors, the dimensions of the style feature vectors correspond to a preset style dimension set, the value of each dimension is the feature value of the style dimension, the Euclidean distance between the two style feature vectors is calculated, the reciprocal of the Euclidean distance is taken as the style similarity of the commodity style attributes, and if the Euclidean distance is zero, the style similarity is set to a preset maximum value. 6.The method of claim 1, wherein, The method comprises the following steps: All commodities involved in the interaction behavior units in the user behavior sequence are collected, the repeated commodities are removed to form a commodity set, and each commodity in the commodity set is taken as an initial node of the dynamic commodity collaborative network; For any two commodities in the commodity set, if there is at least one adjacent interaction behavior unit in the user behavior sequence that involves the two commodities, the commodity attribute dynamic correlation degree corresponding to the adjacent interaction behavior unit is extracted as the initial weight of the edge connecting the two commodity nodes, and if there are multiple connection edges between the two commodities, the average value of all initial weights is taken as the final edge weight; Each node in the dynamic commodity collaborative network is initialized with a node importance parameter, and the initial node importance parameter is the ratio of the number of interaction behavior units in which the commodity appears in the user behavior sequence to the total number of interaction behavior units; A node importance update period is set, real-time interaction behaviors of the user are collected in each update period, and the real-time interaction behaviors are the interaction behavior units generated by the user in the current update period; The node importance parameter of the corresponding node is updated according to the commodities involved in the real-time interaction behaviors, the node importance parameter is increased by a preset increment value if the interaction type of the real-time interaction behavior is purchase interaction, and the node importance parameter is decreased by a preset decrement value if the interaction type is browsing interaction, the increment value and the decrement value corresponding to the add-to-cart interaction and the collection interaction are between the purchase interaction and the browsing interaction; Periodically normalize the weight of all edges in the dynamic commodity collaborative network, so that the sum of the weight of the out edges of each node is a fixed value. The period of normalization is consistent with the period of updating the node importance. 7.The method of claim 6, wherein, The node importance parameter of the corresponding node is updated according to the real-time interaction behavior. If the interaction type of the real-time interaction behavior is purchase interaction, the node importance parameter increases by a preset increment value. If it is browsing interaction, the node importance parameter decreases by a preset decrement value. The increment value and the decrement value corresponding to the add-to-cart interaction and the collection interaction are between the purchase interaction and the browsing interaction, including: Set the corresponding weight adjustment coefficient for different interaction types. The weight adjustment coefficient of purchase interaction is the first coefficient, the weight adjustment coefficient of add-to-cart interaction is the second coefficient, the weight adjustment coefficient of collection interaction is the third coefficient, and the weight adjustment coefficient of browsing interaction is the fourth coefficient. The first coefficient> The second coefficient> The third coefficient> The fourth coefficient, and the fourth coefficient is negative. Obtain the interaction duration of the real-time interaction behavior, and calculate the ratio of the interaction duration to the average interaction duration of the interaction type as the duration influence factor. The value range of the duration influence factor is a preset interval. Multiply the weight adjustment coefficient and the duration influence factor to obtain the node importance adjustment value. The node importance adjustment value of purchase interaction, add-to-cart interaction and collection interaction is positive, and the node importance adjustment value of browsing interaction is negative. Extract the current node importance parameter of the commodity involved in the real-time interaction behavior in the dynamic commodity collaborative network, and add the current node importance parameter and the node importance adjustment value to obtain the updated node importance parameter. Boundary constraint processing is performed on the updated node importance parameter. If the updated node importance parameter is greater than the preset upper limit value, it is set to the preset upper limit value. If it is less than the preset lower limit value, it is set to the preset lower limit value, so that the node importance parameter is within the preset effective range. 8.The method of claim 1, wherein, The path optimization mining processing is performed in the dynamic commodity collaborative network, and the user potential behavior path set is generated by combining the behavior transition correlation degree and the node importance parameter, including: Determine the commodity involved in the last interaction behavior unit in the user behavior sequence as the starting node of path mining, and extract the node importance parameter of the starting node in the dynamic commodity collaborative network as the path starting weight. Taking the starting node as the starting point, traverse all adjacent nodes directly connected to the starting node in the dynamic commodity collaborative network, take the adjacent nodes as candidate next hop nodes, and collect the node importance parameter of each candidate next hop node and the commodity attribute dynamic correlation degree of the connection edge. For each candidate next hop node, extract the behavior transition correlation degree of the interaction behavior unit corresponding to the starting node in the user behavior sequence, and multiply the behavior transition correlation degree, the node importance parameter of the candidate next hop node and the commodity attribute dynamic correlation degree of the connection edge to obtain the path transition probability. The candidate next hop node with the highest path transfer probability is selected as the second node of the path, the node is taken as a new starting node, the steps of adjacent node traversal, path transfer probability calculation and node selection are repeatedly executed until the path length reaches a preset threshold or the adjacent node cannot be continuously traversed, and a preliminary potential behavior path is generated; The preliminary potential behavior path is subjected to path pruning processing, nodes with a node importance parameter lower than a preset node threshold and corresponding subsequent path segments are removed, and a path with all node importance parameters higher than the preset node threshold is retained as an effective potential behavior path; The above steps are repeatedly executed to generate multiple effective potential behavior paths, so that different potential behavior paths contain different combinations of product attribute associated features, and a user potential behavior path set is formed. 9.The method of claim 8, wherein, The preliminary potential behavior path is subjected to path pruning processing, nodes with a node importance parameter lower than a preset node threshold and corresponding subsequent path segments are removed, and a path with all node importance parameters higher than the preset node threshold is retained as an effective potential behavior path, including: Each node in the preliminary potential behavior path is traversed, and the node importance parameter of the node is extracted in sequence according to the path order; The node importance parameter of each node is compared with the preset node threshold, and the first node with a node importance parameter lower than the preset node threshold is marked as a pruning node; If there is no pruning node in the preliminary potential behavior path, the path is directly taken as an effective potential behavior path; If there is a pruning node, the pruning node and all path segments after the node are removed, and the path segment before the pruning node is retained as a candidate pruning path; The path length of the candidate pruning path is calculated, if the path length is less than a preset minimum path length, the candidate pruning path is discarded, otherwise the candidate pruning path is taken as an effective potential behavior path; The above pruning processing steps are repeatedly executed on all preliminary potential behavior paths, and all effective potential behavior paths are collected to form an intermediate path set, so that the node importance parameters of the paths in the intermediate path set meet the preset node threshold requirements. 10.A commodity personalized recommendation system based on user behavior data analysis, characterized in that, The processor and the memory are connected, the memory is used to store programs, instructions or codes, and the processor is used to execute the programs, instructions or codes in the memory to realize the product individualization recommendation method based on user behavior data analysis in any one of claims 1-9.
Citation Information
Patent Citations
Customized recommendation method based on analysis of purchase user behaviors
CN106991592A
Commodity recommendation method based on cooperation of multi-granularity attribute set and neighbor attention
CN116108284A
E-commerce platform commodity recommendation method and system based on user preference analysis
CN119398864A
Personalized commodity recommendation optimization method based on deep learning
CN120374226A
Cited By
User portrait generation method and system based on big data
CN121092969A
Customized furniture marketing and ordering transaction system fused with AR scene
CN121146875A
User intention determination method and device, electronic equipment and storage medium
CN121388284A
Recommendation method and system based on user tag
CN121391429A
E-commerce user re-purchase behavior prediction and accurate reaching method and system fused with time sequence attention mechanism
CN122022900A