Major curve optimization method and device based on feature constraint

By introducing data attribute feature constraint coefficients and adaptive rotation axis direction in the main curve method, the problems of low fitting accuracy and unstable rotation axis in the traditional method are solved, and efficient and accurate high-dimensional nonlinear data analysis is achieved.

CN120234546AActive Publication Date: 2025-07-01UESTC (SHENZHEN) ADVANCED RES INST +1
View PDF 3 Cites 0 Cited by

Patent Information

Application Number
CN202510703976.3
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-05-29
Publication Date
2025-07-01
Estimated Expiration
2045-05-29

AI Technical Summary

Technical Problem

When processing high-dimensional nonlinear data, the existing main curve method ignores the specific attribute characteristics of the data, resulting in low fitting accuracy and unstable rotation axis selection, which affects the convergence of the optimization process and the global consistency of the fitting results.

Method used

By introducing data attribute feature constraint coefficients, combining Riemann distance, the data point contribution factor of high-dimensional nonlinear data is optimized, the initial main curve is constructed, and the adaptive rotation axis direction is determined by minimizing the target optimization error, which solves the problem of ignoring data attribute features in traditional methods.

Benefits of technology

The main curve fitting accuracy and adaptability of high-dimensional nonlinear data is improved, the calculation complexity is reduced, and the global consistency and calculation efficiency of the rotation representation are ensured.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120234546A_ABST
    Figure CN120234546A_ABST
Patent Text Reader

Abstract

The invention is suitable for the technical field of data processing, and provides a main curve optimization method and device based on feature constraint, and the method comprises the steps: obtaining to-be-analyzed high-dimensional nonlinear data, and mapping the to-be-analyzed high-dimensional nonlinear data to a target Riemannian manifold; the high-dimensional nonlinear data has specific attributes; constructing an initial main curve to be optimized through the data points of the target Riemannian manifold; determining a target projection point of the target data point projected to the target position; determining a joint contribution factor of the target data point according to the spatial constraint coefficient and the at least one data attribute feature constraint coefficient; and updating the to-be-optimized main curve by minimizing the target optimization error to obtain a final target main curve. According to the method, the data attribute feature constraint coefficient is introduced on the basis of the spatial constraint coefficient to jointly determine the joint contribution factor of the data points of the high-dimensional nonlinear data, so that the problem of fitting errors caused by neglecting data attribute features in a traditional main curve method is solved, and the main curve fitting precision is improved.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present application relates to the technical field of data processing, and in particular, to a method and device for optimizing a principal curve based on feature constraints. Background Art

[0002] With the rapid development of artificial intelligence and data science, the application of high-dimensional non-linear data in many fields is becoming increasingly widespread. Such data not only contains complex spatial structures, but may also have specific attribute features, such as time series characteristics, density distribution characteristics, or speed change characteristics, etc. For example, human motion data, as typical high-dimensional non-linear data, not only contains complex spatial structures, but also has strong time series characteristics and three-dimensional rotation information. Accurately and effectively analyzing and modeling high-dimensional non-linear data is of great significance for promoting the development of related fields.

[0003] Currently, the principal curve method, as an important non-linear dimensionality reduction technique, has been widely studied and applied. This method captures the main trends and potential structures of data by fitting a smooth curve on the data manifold, providing a basis for subsequent analysis and modeling. However, the existing principal curve methods face two key technical problems when dealing with high-dimensional non-linear data with specific attributes: First, the feature misalignment problem. The existing principal curve methods only rely on the Riemannian distance to optimize the curve during the update process, completely ignoring the specific attribute features that the data may have. For example, when dealing with data with strong time series characteristics (such as human motion data), if the traditional method treats the data points as an unordered set and ignores the time correlation between data points, it will lead to the destruction of the time structure of the data during the fitting process, making the principal curve unable to accurately capture the true evolution process of the data. For example, when the start stage of an action is spatially similar to the end stage of another action, the existing principal curve methods may wrongly confuse these stages, resulting in an incoherent fitting result; or for periodic motions, it is unable to effectively distinguish similar configurations in adjacent periods, often wrongly classifying similar points in different periods into the same fitting group, resulting in the destruction of the periodic structure, thus seriously affecting the fitting accuracy of the principal curve.

[0004] Second, when dealing with data containing rotation information, there is a problem of unstable rotation axis selection. Traditional methods usually use predefined fixed rotation axes (such as the X, Y, or Z axes) to process rotation data. This method ignores the geometric characteristics of the data distribution and is prone to direction instability when dealing with equivalent rotations (for example, rotating 90° to the left is equivalent to rotating 270° to the right mathematically), which in turn affects the convergence of the optimization process and the global consistency of the fitting result. Especially when there are measurement errors or depth estimation biases in the rotation data, this instability will be further amplified, resulting in a significant reduction in the fitting quality of the principal curve and introducing unnecessary computational complexity. Summary of the Invention

[0005] The embodiments of the present application provide a method, device, electronic device and medium for optimizing a principal curve based on feature constraints, which can solve the problem of low fitting accuracy of traditional principal curve methods. The present application provides a method, device and medium for optimizing a principal curve based on feature constraints.

[0006] In a first aspect, the embodiments of the present application provide a method for optimizing a principal curve based on feature constraints, including: Obtain high-dimensional non-linear data to be analyzed and map the high-dimensional non-linear data to a target Riemannian manifold; the high-dimensional non-linear data has specific attributes; Construct an initial principal curve to be optimized through the data points of the target Riemannian manifold; Determine a target projection point where a target data point is projected to a target position; the target data point is any data point of the target Riemannian manifold, and the target position is any position of the principal curve to be optimized; Determine a joint contribution factor of the target data point according to a spatial constraint coefficient and at least one data attribute feature constraint coefficient; wherein, the spatial constraint coefficient is determined based on the Riemannian distance between the target projection point and a target curve point, the data attribute feature constraint coefficient is determined based on the specific attribute, and the target curve point is the point on the principal curve to be optimized at the target position; Update the principal curve to be optimized by minimizing a target optimization error to obtain a final target principal curve; wherein, the target optimization error is weighted by the joint contribution factor.

[0007] The beneficial effects of the embodiments of the present application compared with the prior art are: By introducing a data attribute feature constraint coefficient on the basis of the existing spatial constraint coefficient to jointly determine the joint contribution factor of the data points of the high-dimensional non-linear data, the subsequent update of the principal curve is not only affected by the Riemannian distance-related coefficient, but also affected by the data attribute feature-related coefficient, so as to be able to maintain the specific attribute structure of the data, solve the fitting error problem caused by the traditional principal curve method relying only on spatial distance and ignoring data attribute features, and improve the principal curve fitting accuracy and adaptability of high-dimensional non-linear data.

[0008] In a possible implementation manner of the first aspect, the specific attributes include a timing attribute and a distribution attribute, the timing attribute corresponds to at least a time feature, and the distribution attribute corresponds to at least a density feature.

[0009] In the above solution, different data attribute feature constraints can be flexibly selected according to different specific attributes of the data, which makes this method have wide applicability and can meet the analysis requirements of various complex data. For time series data with strong temporal characteristics, time feature constraints can be selected; for distributed data with uneven distribution, density feature constraints can be selected.

[0010] In a possible implementation manner of the first aspect, when the high-dimensional non-linear data is human motion data, the steps of obtaining the high-dimensional non-linear data to be analyzed and mapping the high-dimensional non-linear data to the target Riemannian manifold include: Obtain the joint point information and joint limb information of the human skeleton; Obtain the rotation data of two adjacent joint points at the same moment from the human motion data, and convert the rotation data into the rotation matrix of the joint limb; Determine all the rotation matrices at the same moment as a set of rotation matrix sequences; among them, multiple sets of rotation matrix sequences at different moments have sequentiality in time; Map the multiple sets of rotation matrix sequences to the target Riemannian manifold.

[0011] In the above solution, when processing human motion data, the human motion is described by using the human skeleton information and the rotation matrix. The rotation matrix sequence can not only accurately and intuitively describe the posture of the joint limb, but also describe the dynamic changes of the joint limb at different moments, and can avoid the inherent singularity problem of representation methods such as Euler angles, thus comprehensively and accurately reflecting the geometric characteristics of human motion. In addition, by emphasizing the sequentiality of the data in time during the processing, the temporal attribute of the data is increased, laying a foundation for introducing constraints on the time characteristics for the subsequent update of the principal curve.

[0012] In a possible implementation manner of the first aspect, the step of constructing the principal curve to be optimized through the data points of the target Riemannian manifold includes: Divide the data points on the Riemannian manifold into multiple data groups; Respectively determine the mean points of the data groups based on Riemannian geometry; Generate intermediate points by geodesic interpolation between two adjacent mean points; Fit the principal curve to be optimized according to the mean points and the intermediate points.

[0013] In the above solution, the Riemann mean point obtained by calculation is used as the representative node of the principal curve to form the initial trajectory, thereby preserving the true characteristics of the data, reducing the dependence on hyperparameters, enhancing the generalization ability on different data types, and then smoothing the initial trajectory through geodesic interpolation to generate a continuous and smooth principal curve. Compared with random selection or linear interpolation, this principal curve has significantly lower fitting error, thereby improving the fitting efficiency and fitting accuracy in the subsequent principal curve update process.

[0014] In a possible implementation manner of the first aspect, when the data point corresponds to a rotation matrix, the step of respectively determining the mean point of the data grouping based on Riemannian geometry includes: Using logarithmic mapping to map each rotation matrix in the data grouping to a rotation vector in the Lie algebra; Calculating the average vector based on the rotation vector; Using exponential mapping to map the average vector back to the Lie group to obtain an initial candidate mean point; Calculating the Riemannian distance from each rotation matrix in the data grouping to the candidate mean point; Updating the candidate mean point by minimizing the objective function to obtain the final mean point of the data grouping, and the minimizing objective function is: ; Wherein, represents the mean point, R represents the candidate mean point, represents the rotation matrix in the data grouping, SO(3) represents the three-dimensional special rotation group, represents the Riemannian distance from the rotation matrix to the candidate mean point.

[0015] In the above solution, by dynamically selecting the data point closest to the center as the initial candidate mean point, the overall distribution of the data can be more accurately reflected, thereby reducing the subsequent number of iterations and obtaining a result closer to the true mean faster and more accurately. For high-dimensional non-linear data, by utilizing the correspondence between Lie groups and Lie algebras, the optimization problem on the high-dimensional manifold space is reduced to an optimization on a low-dimensional space, greatly improving the calculation efficiency. Especially for data processing containing rotation information, the computational complexity is reduced from O(n³) to O(n), which is more suitable for real-time applications or computationally resource-constrained environments.

[0016] In a possible implementation manner of the first aspect, the joint contribution factor of the target data point is determined by the following formula: ; Wherein, t represents the position parameter of the principal curve, n represents the identification parameter of the data point, represents the joint contribution factor, represents the spatial constraint coefficient, represents the constraint coefficient of the data attribute feature, represents the smoothing kernel function.

[0017] In the above solution, a joint smoothing kernel function that simultaneously considers spatial factors and data attribute features is designed to calculate the joint contribution factor. This way of introducing data attribute feature constraints is easy to apply to optimizing the principal curve fitting process of various different types of data.

[0018] In a possible implementation manner of the first aspect, when the specific attribute is a time series attribute, the joint contribution factor of the target data point is determined by the following formula: ; where, represents the time feature constraint coefficient, represents the curve point at parameter t on the principal curve, represents the data point the projection point of the data point at parameter t on the principal curve, represents the Riemannian distance between the projection point and the curve point, represents the spatial scale factor, represents the time scale factor, represents the time deviation between the data point and the curve point, represents the time corresponding to the curve point with parameter t on the principal curve, represents the data point the corresponding time.

[0019] In a possible implementation manner of the first aspect, when the high-dimensional non-linear data contains rotation information, the principal curve optimization method based on feature constraints further includes determining the adaptive rotation axis direction of the high-dimensional non-linear data. The steps of determining the adaptive rotation axis direction of the high-dimensional non-linear data include: Converting the high-dimensional non-linear data into at least one rotation matrix dataset; the rotation matrix dataset includes multiple rotation matrices; for any one rotation matrix dataset, determining the mean matrix of the rotation matrix dataset based on Riemannian geometry; Obtaining the deviation vectors between the mean matrix and each rotation matrix in the rotation matrix dataset; Constructing a covariance matrix based on the deviation vectors; Performing eigenvalue decomposition on the covariance matrix to obtain eigenvalues and corresponding eigenvectors; the eigenvalues represent the amount of change in the corresponding eigenvector direction, and the eigenvectors are used to represent the change direction of the rotation matrix dataset; Selecting the eigenvector corresponding to the smallest eigenvalue and determining it as the rotation axis direction of the rotation matrix dataset; Determine the adaptive rotation axis direction of the high-dimensional non-linear data based on all different rotation matrix datasets.

[0020] In the above solution, for the high-dimensional non-linear data containing rotation information, the determination of the adaptive rotation axis is carried out, getting rid of the limitations of the traditional fixed-axis method, and being able to automatically determine the optimal rotation axis according to the changing characteristics in the data distribution. The direction corresponding to the minimum eigenvalue in the covariance matrix represents the direction with the least change in the dataset. By selecting this direction as the rotation axis, it can effectively avoid the mirror ambiguity and direction instability problems caused by equivalent rotation, ensuring the global consistency and reliability of the rotation representation. In addition, for the rotation matrix, by constraining a suitable rotation axis in the dataset, the high-dimensional optimization problem can be reduced to an optimization problem in a low-dimensional space, greatly reducing the computational cost and improving the computational efficiency.

[0021] In a second aspect, an embodiment of the present application provides a principal curve optimization device based on feature constraints, including: A data acquisition module, configured to acquire the high-dimensional non-linear data to be analyzed and map the high-dimensional non-linear data to a target Riemannian manifold; the high-dimensional non-linear data has specific attributes; A curve construction module, configured to construct an initial principal curve to be optimized through the data points of the target Riemannian manifold; A projection module, configured to determine a target projection point for projecting a target data point to a target position; the target data point is any data point of the target Riemannian manifold, and the target position is any position of the principal curve to be optimized; A contribution calculation module, configured to determine a joint contribution factor of the target data point according to a spatial constraint coefficient and at least one data attribute feature constraint coefficient; wherein, the spatial constraint coefficient is determined based on the Riemannian distance between the target projection point and the target curve point, the data attribute feature constraint coefficient is determined based on the specific attribute, and the target curve point is the point on the principal curve to be optimized at the target position; A curve update module, configured to update the principal curve to be optimized by minimizing the target optimization error to obtain a final target principal curve; wherein, the target optimization error is weighted by the joint contribution factor.

[0022] In a third aspect, an embodiment of the present application provides an electronic device, including a memory, a processor, and a computer program stored in the memory and executable on the processor. When the processor executes the computer program, it implements the principal curve optimization method according to any one of the first aspects above.

[0023] Fourthly, an embodiment of the present application provides a computer-readable storage medium storing a computer program, and when the computer program is executed by a processor, it implements the method for optimizing the principal curve based on feature constraints described in any one of the above first aspects.

[0024] Fifthly, an embodiment of the present application provides a computer program product, and when the computer program product runs on a terminal device, it causes the terminal device to execute the method for optimizing the principal curve based on feature constraints described in any one of the above first aspects.

[0025] It can be understood that the beneficial effects of the above second to fifth aspects can be referred to the relevant descriptions in the above first aspect, and will not be elaborated here. Description of the Drawings

[0026] Figure 1 is a flowchart of the method for optimizing the principal curve based on feature constraints provided by an embodiment of the present application; Figure 2 is a schematic diagram of the principal curve of the prior art; Figure 3 is a schematic diagram of the principal curve of an embodiment of the present application; Figure 4 is a schematic diagram of the principal curve of another embodiment of the present application; Figure 5 is a schematic diagram of the human skeleton provided by an embodiment of the present application; Figure 6 is a flowchart of determining the adaptive rotation axis provided by an embodiment of the present application; Figure 7 is a schematic diagram of the structure of the device for optimizing the principal curve based on feature constraints provided by an embodiment of the present application; Figure 8 is a schematic diagram of the structure of the electronic device provided by an embodiment of the present application. Detailed Embodiments

[0027] It should be noted that, without conflict, the embodiments in the present application and the features in the embodiments can be combined with each other.

[0028] In the following description, for the purpose of illustration rather than limitation, specific details such as specific system structures and technologies are proposed to thoroughly understand the embodiments of the present application. However, those skilled in the art should clearly understand that the present application can also be implemented in other embodiments without these specific details. In other cases, the detailed descriptions of well-known systems, devices, circuits, and methods are omitted to avoid unnecessary details from interfering with the description of the present application.

[0029] It should be understood that when used in the specification of this application and the appended claims, the term "comprising" indicates the presence of the described features, wholes, steps, operations, elements, and / or components, but does not exclude the presence or addition of one or more other features, wholes, steps, operations, elements, components, and / or their combinations.

[0030] It should also be understood that the term "and / or" used in the specification of this application and the appended claims refers to any combination and all possible combinations of one or more of the associated listed items, and includes these combinations.

[0031] As used in the specification of this application and the appended claims, the term "if" can be interpreted as "when", "once", "in response to determining", or "in response to detecting" depending on the context. Similarly, the phrase "if determined" or "if [the described condition or event] is detected" can be interpreted as meaning "once determined", "in response to determining", "once [the described condition or event] is detected", or "in response to detecting [the described condition or event]" depending on the context.

[0032] In addition, in the description of the specification of this application and the appended claims, the terms "first", "second", "third", etc. are only used for distinguishing descriptions and should not be construed as indicating or implying relative importance.

[0033] See Figure 1 , the embodiment of this application provides a schematic flowchart of a method for optimizing a principal curve based on feature constraints. By way of example and not limitation, the method may include the following steps: S11. Obtain high-dimensional non-linear data to be analyzed and map the high-dimensional non-linear data to a target Riemannian manifold.

[0034] Among them, the high-dimensional non-linear data has specific attributes. In a possible implementation, the specific attributes include a temporal attribute and a distribution attribute. Among them, the temporal attribute corresponds at least to a time feature, and the distribution attribute corresponds at least to a density feature. Of course, this embodiment is not limited to the above two attributes and also includes a dynamic attribute, which corresponds at least to a speed feature.

[0035] In this embodiment, each data point of the target Riemannian manifold represents the geometric characteristics of a data point. In a possible implementation, when processing data containing rotation information (such as human motion data), the target Riemannian manifold is SO(3). SO(3) can be used both as a three-dimensional special rotation group and as a three-dimensional Riemannian manifold. Its group characteristics reflect the algebraic structure of the rotation relationship, and its manifold characteristics reveal its geometric topological properties, providing a basis for the construction and optimization of the subsequent principal curve.

[0036] Of course, the target Riemannian manifold in this embodiment is not limited to SO(3). An appropriate Riemannian manifold can be selected according to the data characteristics of high-dimensional non-linear data, application requirements, and the geometric properties of the manifold. For example, for electroencephalogram signal analysis data, the covariance matrix between data points is relatively important, and the SPD manifold (symmetric positive definite matrix manifold) can be selected; for another example, for word vector data, to capture the hierarchical semantic relationship of vocabulary, the hyperbolic space can be selected, and so on.

[0037] S12. Construct an initial main curve to be optimized through the data points of the target Riemannian manifold.

[0038] S13. Determine the target projection point where the target data point is projected to the target position.

[0039] Among them, the target data point is any data point of the target Riemannian manifold, and the target position is any position of the main curve to be optimized.

[0040] S14. Determine the joint contribution factor of the target data point according to the spatial constraint coefficient and at least one data attribute feature constraint coefficient.

[0041] Among them, the spatial constraint coefficient is determined based on the Riemannian distance between the target projection point and the target curve point, the data attribute feature constraint coefficient is determined based on a specific attribute, and the target curve point is the point on the main curve to be optimized at the target position.

[0042] S15. Update the main curve to be optimized by minimizing the target optimization error to obtain the final target main curve.

[0043] Among them, the target optimization error is weighted by the joint contribution factor. Specifically, when the target optimization error reaches the minimum value, the main curve to be optimized is determined as the target main curve; otherwise, based on the projection point, fit the main curve to be optimized for the next iteration optimization, and then return to execute step S3.

[0044] A method for optimizing the main curve based on feature constraints provided by this embodiment jointly determines the joint contribution factor of the data points of high-dimensional non-linear data by introducing a data attribute feature constraint coefficient on the basis of the existing spatial constraint coefficient, so that the subsequent update of the main curve is affected not only by the Riemannian distance correlation coefficient but also by the data attribute feature correlation coefficient, thereby being able to maintain the specific attribute structure of the data, solving the fitting error problem caused by the traditional main curve method relying only on spatial distance and ignoring data attribute features, and improving the fitting accuracy and adaptability of the main curve of high-dimensional non-linear data.

[0045] In one implementation manner of step S4, the joint contribution factor of the target data point is determined by the following formula: ; Among them, t represents the position parameter of the main curve, that is, the independent variable in the parametric representation of the main curve, which is used to locate points on the curve, and n represents the identification parameter of the data point. represents the joint contribution factor. represents the spatial constraint coefficient. represents the data attribute feature constraint coefficient. represents the smoothing kernel function.

[0046] The spatial distance constraint coefficient c(t, n) is calculated by the following formula: ; Among them, represents the Riemannian distance between the projection point and the curve point. represents the curve point at parameter t on the main curve (that is, at position t). represents the data point. represents the data point at the projection position on the main curve. represents the data point at the projection point at parameter t on the main curve. represents the spatial scale factor. indicates that this formula holds for all parameters n and parameter t.

[0047] A method for optimizing the main curve based on feature constraints provided in this embodiment calculates the joint contribution factor by designing a joint smoothing kernel function that simultaneously considers spatial factors and data attribute features. This way of introducing data attribute feature constraints is easy to apply to the process of optimizing the main curve fitting of various different types of data.

[0048] In a possible implementation, when the specific attribute is a time series attribute, step S14 includes determining the joint contribution factor of the target data point by the following formula: ; Among them, represents the time feature constraint coefficient. represents the time deviation between the data point and the curve point. Specifically, represents the time corresponding to the curve point with parameter t on the main curve. represents the data point at the value in the time dimension, that is, the data point corresponding time. represents the time scale factor. When the time deviation exceeds the preset threshold, even if the data point is close to the curve point in space, its joint contribution factor will be significantly reduced, thereby maintaining the temporal consistency of curve update.

[0049] The following refers toFigure 2 and Figure 3 are illustrated by examples. As Figure 2 shown, the curve update in the existing principal curve fitting method only depends on the Riemannian distance. Under this mechanism, the update of data points is very likely to be affected by data points that are far in time distance but close in spatial distance. For example Figure 2 the third point on the curve in wrongly classifies the data point into the same fitting group, resulting in the updated point deviating from the correct time evolution trajectory.

[0050] The principal curve is updated by jointly considering spatial constraints and time constraints simultaneously. As Figure 3 shown, since the data point exceeds the time constraint, although it is spatially close to the data point , it is excluded from the update range when updating the data point , thus classifying the data point and the data point into the same fitting group. Therefore, Figure 3 the updated data point in

[0051] In another possible implementation, when the specific attribute is a distribution attribute, step S14 includes calculating the joint contribution factor by the following formula: ; wherein, represents the density constraint coefficient, represents the number of neighboring data points of each curve point within a certain range, represents the density scale factor.

[0052] As Figure 4 shown, under the constraint of density, the data points in the high-density region have higher weights, and the principal curve actively focuses on the data trend with higher density, so that the principal curve better reflects the main change direction of the data.

[0053] Optionally, an implementation of step S14 may include: determining the joint contribution factor of the target data point according to at least two of the spatial constraint coefficient, the time constraint coefficient, and the density constraint coefficient, which are data attribute feature constraint coefficients. This multi-constraint combination method makes the method have a wider applicability and can flexibly select and combine different types of constraints according to different characteristics of the data.

[0054] It should be noted that in addition to the three data attribute feature constraint coefficients shown in this embodiment, step S14 can also set other feature constraint coefficients based on requirements, such as speed constraint coefficients or distribution constraint coefficients, etc., to adapt to the characteristics and analysis requirements of different types of data. Specifically, for data with strong temporal characteristics, time constraints can be selected; for data with uneven distribution, density constraints can be selected; for data sensitive to the rate of change, speed constraints can be selected. This flexible constraint selection mechanism enables this method to adapt to the analysis requirements of various complex data, and when the data has a combination of multiple specific attributes, multiple data attribute feature constraint coefficients can be introduced.

[0055] Optionally, one implementation manner of step S15 may include: Based on the joint contribution factor of each data point and the Riemannian distance between the corresponding projection point and curve point calculate the target optimization error , specifically: ; The target optimization error represents the sum of the weighted squared distances between all curve points on the principal curve and the projection points of the data points on the principal curve. When this value reaches the minimum, this principal curve is the finally determined principal curve. Before that, the shape of the principal curve can be continuously adjusted according to the projection points. The specific formula is as follows: ; As an example rather than a limitation, when the high-dimensional non-linear data is human motion data, the high-dimensional non-linear data can be converted into multiple sets of rotation matrix series. Optionally, one implementation manner of step S11 may include: S111. Obtain the joint point information and joint limb information of the human skeleton.

[0056] S112. Obtain the rotation data of two adjacent joint points at the same moment from the human motion data, and convert the rotation data into the rotation matrix of the joint limb.

[0057] S113. Aggregate the rotation matrices between all limb pairs at the same moment into a set of rotation matrix sequences.

[0058] Among them, multiple sets of rotation matrix sequences at different moments have sequentiality in time.

[0059] S114. Map multiple sets of rotation matrix sequences to the target Riemannian manifold.

[0060] In a possible implementation manner, the joint point information is defined as , and the joint limb information is defined as , where n represents the subscript parameter of the joint point, Denote the nth joint point, and m represents the subscript parameter of the joint limb. Denote the mth joint limb; define and as two adjacent joint points. Meanwhile, and can also be regarded as the two end points of the joint limb . Represent the rotation data of and at time s as the rotation matrix of the joint limb . The rotation matrix R is a 3×3 orthogonal matrix that satisfies and , where is the transpose matrix of R, I is the 3×3 identity matrix, and det(R)=1 indicates that the determinant of the rotation matrix is 1. Collect all the rotation matrices to obtain the rotation matrix sequence .

[0061] The following is an example with reference to Figure 5 . Figure 5 shows that there are 20 joint points in the skeleton model, connected by 19 joint limbs. The rotation data of the human body can be regarded as the rotation matrix sequence . In a possible implementation, when the human body movement is limited to some joint limbs, a rotation matrix sequence can be established only for some joint limbs. For example, Figure 5 only shows the right hand movement. Figure 5 In , the left skeleton model is equivalent to the initialized right hand movement, and the right skeleton model is equivalent to the changed right hand movement. The human body motion data can be represented only by the rotation matrix sequence

[0062] . It should be noted that all the rotation matrix sequences are arranged in chronological order, so that the human body motion data has the specific attribute of temporal characteristics.

[0063] A method for optimizing the principal curve based on feature constraints provided in this embodiment, when processing the human body motion data, uses the human skeleton information and the rotation matrix to describe the human body motion. The rotation matrix sequence can not only accurately and intuitively describe the posture of the joint limb, but also describe the dynamic changes of the joint limb at different times, and can avoid the inherent singularity problems of representation methods such as Euler angles, thus comprehensively and accurately reflecting the geometric characteristics of the human body motion. In addition, by emphasizing the sequentiality of the data in time during the processing, the temporal attribute of the data is increased, laying a foundation for introducing constraints on the time characteristics in the subsequent update of the principal curve.

[0064] Optionally, an implementation of step S12 may include: S121. Divide the data points on the Riemannian manifold into multiple data groups.

[0065] In a possible implementation, for data with temporal attributes, it can be divided according to time steps. For example, in human motion data, a sequence of rotation matrices at a certain moment is a data group. Of course, other methods can also be used to divide the data.

[0066] S122. Based on Riemannian geometry, determine the mean points of the data groups respectively.

[0067] S123. Between two adjacent mean points, generate intermediate points using geodesic interpolation.

[0068] S124. Fit the main curve to be optimized according to the mean points and intermediate points.

[0069] A method for optimizing the main curve based on feature constraints provided in this embodiment uses the calculated Riemannian mean points as the representative nodes of the main curve to form an initial trajectory, thereby retaining the true features of the data, reducing the dependence on hyperparameters, enhancing the generalization ability on different data types, and then smoothing the initial trajectory through geodesic interpolation to generate a continuous and smooth main curve. Compared with random selection or linear interpolation, this main curve has significantly lower fitting errors, thereby improving the fitting efficiency and fitting accuracy in the subsequent main curve update process.

[0070] As an example rather than a limitation, when the data points correspond to rotation matrices, the data group is represented as a series of rotation matrices composed of multiple rotation matrices. Optionally, an implementation of step S121 may include: S1211. Use logarithmic mapping to map each rotation matrix in the data group to a rotation vector in the Lie algebra.

[0071] Specifically, given a sequence of rotation matrices , representing all the rotation matrices in the data group, according to the properties of SO(3), map each rotation matrix to the Lie algebra , is a vector space composed of three-dimensional skew-symmetric matrices , and the skew-symmetric matrix corresponds one-to-one with the rotation vector . The formula for logarithmic mapping is: ; In a possible implementation, the rotation vector is given by the product of the rotation angle and the unit rotation axis : ;

[0072] In one possible implementation, the rotation angle is related to the trace of the rotation matrix . The trace of the rotation matrix is the sum of the elements on the main diagonal of the matrix. The rotation angle and the rotation axis can be calculated using the following formula. Specifically: ; ; ; ; S1212. Calculate the average vector based on the rotation vector.

[0073] In one possible implementation, the average vector can be calculated using the following formula: ; S1213. Use the exponential map to map the average vector back into the Lie group to obtain the initial candidate mean point.

[0074] In one possible implementation, the exponential map can map the average vector back to the rotation matrix (i.e., the candidate mean point). Specifically: ; where is the skew-symmetric matrix of the average vector , represents the rotation angle, and can be expressed as: ; ;

[0075] S1214. Calculate the Riemannian distance from each rotation matrix in the data group to the candidate mean point.

[0076] In one possible implementation, the Riemannian distance can be calculated as follows. Specifically: ; where represents the rotation matrix in the data group, represents the candidate mean point, represents the Riemannian distance from the rotation matrix to the candidate mean point, represents the Frobenius norm, which is the square root of the sum of the squares of all elements of the matrix.

[0077] The calculation of the Frobenius norm can be to take the square root after summing the squares of all elements of the matrix. For example, for a matrix , its Frobenius norm is .

[0078] S1215. Update the candidate mean points by minimizing the objective function to obtain the mean points of the final rotation matrix sequence.

[0079] Among them, the objective function to be minimized is: ; Among them, represents the finally determined mean point. By continuously adjusting , the value of on the right side of the equation reaches the minimum value. At this time, the determined is the finally determined mean point .

[0080] A method for optimizing the principal curve based on feature constraints provided by this embodiment can more accurately reflect the overall distribution of data by dynamically selecting the data points closest to the center as the initial candidate mean points, thereby reducing the subsequent number of iterations and obtaining a result closer to the true mean faster and more accurately. For high-dimensional non-linear data, by using the correspondence between Lie groups and Lie algebras, the optimization problem on the high-dimensional manifold space is reduced to an optimization on a low-dimensional space, greatly improving the computational efficiency. Especially for data processing containing rotation information, the computational complexity is reduced from O(n³) to O(n), making it more suitable for real-time applications or computing environments with limited resources.

[0081] Optionally, a specific implementation of step S122 specifically includes: On the SO(3) manifold, each rotation matrix corresponds to a tangent space . Assuming two mean matrices and , calculate the tangent vector from to based on the logarithmic mapping. Specifically: ; Generate an interpolation path (i.e., geodesic from to ) through the following geodesic formula:

[0082] Among them, by selecting different t values (e.g., t = 0, 0.1, 0.2,..., 1), a series of interpolation points (i.e., intermediate points) can be generated.

[0083] Optionally, an implementation manner of step S123 specifically includes: using the above-mentioned mean point and midpoint as curve points on the main curve, and generating an initial main curve to be optimized by connecting the curve points.

[0084] When processing rotational data, equivalent rotations on the SO(3) manifold (e.g., 90° left rotation and 270° right rotation) may lead to unstable mean calculations, affecting the optimization process of the main curve. Especially in human motion data, the instability of joint rotation representations directly affects the accuracy and consistency of motion modeling.

[0085] The reference to a fixed rotation axis can solve the rotation instability problem and also significantly reduce the computational amount. After the data is projected onto the fixed rotation axis, the rotation matrix can be decomposed into the rotation angle around this axis, thus simplifying the data representation from high-dimensional rotation to scalar angle representation. However, the choice of the rotation axis largely determines the quality of data dimensionality reduction. When different actions involve different rotation trends, a fixed rotation axis may lead to the problem of local information loss. For example, in actions mainly moving along the Z-axis such as "jumping" and "squatting", fixing the Z-axis helps to maximize the retention of their key features. For actions involving multi-directional rotations (such as "waving" or "moving an object"), the selection of a single axis may lead to the loss of some motion information. Therefore, it is necessary to select an appropriate rotation axis according to different types of human motion data.

[0086] In a possible implementation manner, when the high-dimensional non-linear data contains rotation information, this main curve optimization method based on feature constraints further includes determining the adaptive rotation axis direction of the high-dimensional non-linear data, such as Figure 6 As shown, the steps of determining the adaptive rotation axis direction of the high-dimensional non-linear data include: S21: Convert the high-dimensional non-linear data into at least one rotation matrix data set.

[0087] Among them, the rotation matrix data set includes multiple rotation matrices. For complex high-dimensional non-linear data, a pre-partitioning process can be performed in advance. Each rotation matrix data set is represented as , represents the rotation matrix, and SO(3) represents the three-dimensional special rotation group.

[0088] S22: For any one rotation matrix data set, determine the mean matrix of the rotation matrix data set based on Riemannian geometry.

[0089] In this embodiment, the implementation manner of step S22 is basically the same as that of step S121. The rotation matrix data set corresponds to the rotation matrix sequence, and the mean matrix corresponds to the mean point, which will not be elaborated here.

[0090] S23. Obtain the deviation vectors of each rotation matrix in the mean matrix and the rotation matrix dataset.

[0091] In a possible implementation, through logarithmic mapping , each rotation matrix can be converted into an anti-symmetric matrix in Lie algebra , establishing a close connection between the Lie group and its Lie algebra. The Lie algebra consists of all anti-symmetric matrices. The Lie algebra is the tangent space of the SO(3) manifold (Lie group) at the rotation matrix R (identity element). Then, the deviation of each rotation matrix from the mean mapped to the tangent space through logarithmic mapping is: ; where is the rotation vector in the tangent space, represents the deviation vector of the rotation matrix relative to the mean matrix, that is, it represents the direction of the rotation deviation on the manifold.

[0092] S24. Construct a covariance matrix based on the deviation vectors.

[0093] Among them, the covariance matrix C describes the variations of the rotation matrix dataset in different directions. Optionally, the covariance matrix C of the rotation matrix series is: ; S25. Perform eigenvalue decomposition on the covariance matrix to obtain eigenvalues and corresponding eigenvectors.

[0094] Among them, the eigenvalues represent the amount of variation in the directions of the corresponding eigenvectors, and the eigenvectors are used to represent the variation directions of the rotation matrix dataset. Specifically, . Among them is the eigenvalue, representing the variance of the variation along different directions, is the corresponding eigenvector, representing the direction of the variation in the rotation matrix sequence.

[0095] S26. Select the eigenvector corresponding to the smallest eigenvalue and determine it as the rotation axis direction of the rotation matrix dataset.

[0096] Among them, the eigenvector corresponding to the smallest eigenvalue represents the direction with the least variation, that is, the most stable direction, in the rotation matrix dataset. Selecting this direction as the rotation axis can effectively avoid the problem of direction instability caused by equivalent rotation.

[0097] S27. Determine the adaptive rotation axis direction of the high-dimensional non-linear data based on the rotation axis directions of all different rotation matrix datasets.

[0098] A method for optimizing a principal curve based on feature constraints provided in this embodiment determines an adaptive rotation axis for high-dimensional non-linear data containing rotation information, getting rid of the limitations of traditional fixed-axis methods and being able to automatically determine the optimal rotation axis according to the changing characteristics in the data distribution. The direction corresponding to the minimum eigenvalue in the covariance matrix represents the direction with the least change in the dataset. By selecting this direction as the rotation axis, it can effectively avoid the mirror ambiguity and direction instability problems caused by equivalent rotation, ensuring the global consistency and reliability of the rotation representation. In addition, for the rotation matrix, by constraining a suitable rotation axis in the dataset, the high-dimensional optimization problem can be reduced to an optimization problem in a low-dimensional space, greatly reducing the computational cost and improving the computational efficiency.

[0099] It should be understood that the magnitudes of the sequence numbers of the steps in the above embodiments do not mean the order of execution. The order of execution of each process should be determined according to its function and internal logic, and should not constitute any limitation to the implementation process of the embodiments of this application.

[0100] Corresponding to the method for optimizing a principal curve based on feature constraints described in the above embodiments, Figure 7 The structural block diagram of a device for optimizing a principal curve based on feature constraints provided by an embodiment of this application is shown. For the sake of convenience of description, only the parts related to the embodiments of this application are shown.

[0101] Referring to Figure 7 , the device for optimizing a principal curve based on feature constraints includes: A data acquisition module 11, configured to acquire high-dimensional non-linear data to be analyzed and map the high-dimensional non-linear data to a target Riemannian manifold; the high-dimensional non-linear data has specific attributes.

[0102] A curve construction module 12, configured to construct an initial principal curve to be optimized through the data points of the target Riemannian manifold.

[0103] A projection module 13, configured to determine a target projection point where a target data point is projected to a target position. The target data point is any data point of the target Riemannian manifold, and the target position is any position of the principal curve to be optimized.

[0104] A contribution calculation module 14, configured to determine a joint contribution factor of the target data point according to a spatial constraint coefficient and at least one data attribute feature constraint coefficient. Among them, the spatial constraint coefficient is determined based on the Riemannian distance between the target projection point and the target curve point, the data attribute feature constraint coefficient is determined based on the specific attribute, and the target curve point is the point on the principal curve to be optimized at the target position.

[0105] The curve update module 15 is configured to update the main curve to be optimized by minimizing the target optimization error, so as to obtain the final target main curve. The target optimization error is weighted by the joint contribution factor.

[0106] In some embodiments of the present application, the specific attributes include a timing attribute and a distribution attribute. The timing attribute corresponds to at least a time feature, and the distribution attribute corresponds to at least a density feature.

[0107] In some embodiments of the present application, the data acquisition module 11 may be specifically configured to, when the high-dimensional non-linear data is human motion data, acquire the joint point information and joint limb information of the human skeleton; acquire the rotation data of two adjacent joint points at the same moment from the human motion data, and convert the rotation data into a rotation matrix of the joint limb; determine all the rotation matrices at the same moment as a set of rotation matrix sequences; wherein, multiple sets of rotation matrix sequences at different moments are sequential in time; map the multiple sets of rotation matrix sequences to the target Riemannian manifold.

[0108] In some embodiments of the present application, the curve construction module 12 may be specifically configured to divide the data points on the Riemannian manifold into multiple data groups; respectively determine the mean points of the data groups based on Riemannian geometry; generate intermediate points by geodesic interpolation between two adjacent mean points; fit the main curve to be optimized according to the mean points and the intermediate points.

[0109] In some embodiments of the present application, the curve construction module 12 may also be specifically configured to, when the data points correspond to rotation matrices, use logarithmic mapping to map each rotation matrix in the data group to a rotation vector in the Lie algebra; calculate the average vector based on the rotation vectors; use exponential mapping to map the average vector back to the Lie group to obtain an initial candidate mean point; calculate the Riemannian distance from each rotation matrix in the data group to the candidate mean point; update the candidate mean point by minimizing the objective function to obtain the final mean point of the data group.

[0110] In some embodiments of the present application, the main curve optimization device based on feature constraints further includes a rotation axis determination module, which is used to convert the high-dimensional non-linear data into at least one rotation matrix dataset when the high-dimensional non-linear data contains rotation information; for any one rotation matrix dataset, determine the mean matrix of the rotation matrix dataset based on Riemannian geometry; obtain the deviation vectors of the mean matrix and each rotation matrix in the rotation matrix dataset; construct a covariance matrix based on the deviation vectors; perform eigenvalue decomposition on the covariance matrix to obtain eigenvalues and corresponding eigenvectors; the eigenvalues represent the amount of change in the direction of the corresponding eigenvectors, and the eigenvectors are used to represent the change direction of the rotation matrix dataset; select the eigenvector corresponding to the smallest eigenvalue and determine it as the rotation axis direction of the rotation matrix dataset; determine the adaptive rotation axis direction of the high-dimensional non-linear data based on the rotation axis directions of all different rotation matrix datasets.

[0111] It should be noted that for the information interaction, execution process, etc. between the above-mentioned devices / units, since they are based on the same concept as the method embodiments of the present application, their specific functions and the technical effects brought can be specifically referred to the method embodiment part, and will not be elaborated here.

[0112] Those skilled in the art can clearly understand that for the convenience and simplicity of description, only the above-mentioned division of each functional unit and module is used as an example. In practical applications, the above functions can be allocated to different functional units and modules according to needs, that is, the internal structure of the device is divided into different functional units or modules to complete all or part of the functions described above. Each functional unit and module in the embodiment can be integrated into a processing unit, or each unit can exist physically alone, or two or more units can be integrated into one unit. The above integrated unit can be implemented in the form of hardware or in the form of a software functional unit. In addition, the specific names of each functional unit and module are only for the convenience of mutual distinction and do not limit the protection scope of the present application. The specific working process of the units and modules in the above device can refer to the corresponding process in the foregoing method embodiments and will not be elaborated here.

[0113] Figure 8 It is a schematic structural diagram of an electronic device provided in an embodiment of the present application. As Figure 8 shown, the electronic device 2 in the embodiment includes: at least one processor 20 ( Figure 8 only one is shown in the figure), a processor, a memory 21, and a computer program 22 stored in the memory 21 and executable on the at least one processor 20. When the processor 20 executes the computer program 22, it implements the steps in each of the above-mentioned main curve optimization method embodiments based on feature constraints.

[0114] The electronic device 2 can be a computing device such as a desktop computer, a notebook, a handheld computer, and a cloud server. The electronic device 2 may include, but is not limited to, a processor 20 and a memory 21. Those skilled in the art can understand that Figure 8 merely examples of the electronic device 2, which do not constitute a limitation on the electronic device 2, may include more or fewer components than shown in the figure, or combine certain components, or different components. For example, it may also include input / output devices, network access devices, etc.

[0115] The processor 20 can be a central processing unit (CPU), and the processor 20 can also be other general-purpose processors, digital signal processors (DSPs), application specific integrated circuits (ASICs), field-programmable gate arrays (FPGAs), or other programmable logic devices, discrete gate or transistor logic devices, discrete hardware components, etc. The general-purpose processor can be a microprocessor or the processor can also be any conventional processor, etc.

[0116] In some embodiments, the memory 21 can be an internal storage unit of the electronic device 2, such as the hard disk or memory of the electronic device 2. In other embodiments, the memory 21 can also be an external storage device of the electronic device 2, such as a plug-in hard disk, a smart media card (SMC), a secure digital (SD) card, a flash card, etc., equipped on the electronic device 2. Further, the memory 21 can also include both the internal storage unit and the external storage device of the electronic device 2. The memory 21 is used to store an operating system, application programs, a boot loader, data, and other programs, such as the program code of a computer program. The memory 21 can also be used to temporarily store data that has been output or will be output.

[0117] The embodiments of the present application also provide a computer-readable storage medium storing a computer program, and when the computer program is executed by a processor, the steps in the above-mentioned embodiments of the main curve optimization method based on various feature constraints can be implemented.

[0118] The embodiments of the present application provide a computer program product, and when the computer program product runs on a mobile terminal, the mobile terminal is caused to execute the steps in the above-mentioned embodiments of the main curve optimization method based on various feature constraints.

[0119] In the above embodiments, the descriptions of the respective embodiments have their own focuses. For parts not described or recorded in a certain embodiment, reference may be made to the relevant descriptions of other embodiments.

[0120] Those of ordinary skill in the art can realize that the units and algorithm steps of the examples described in combination with the embodiments disclosed herein can be implemented by electronic hardware, or a combination of computer software and electronic hardware. Whether these functions are executed in a hardware or software manner depends on the specific application and design constraints of the technical solution. A professional technician can use different methods to implement the described functions for each specific application, but such implementation should not be considered to exceed the scope of this application.

[0121] In the embodiments provided in this application, it should be understood that the disclosed device / network device and method can be implemented in other ways. For example, the device / network device embodiments described above are only illustrative. For example, the division of the modules or units is only a logical function division. In actual implementation, there may be other division methods. For example, multiple units or components can be combined or integrated into another system, or some features can be ignored or not executed. Another point is that the displayed or discussed coupling or direct coupling or communication connection to each other can be through some interfaces. The indirect coupling or communication connection of the device or unit can be in an electrical, mechanical or other form.

[0122] The above-described embodiments are only used to illustrate the technical solutions of this application, rather than to limit it; although this application has been described in detail with reference to the foregoing embodiments, those of ordinary skill in the art should understand that they can still modify the technical solutions recorded in the foregoing embodiments, or perform equivalent replacements for some of the technical features; and these modifications or replacements do not cause the essence of the corresponding technical solutions to deviate from the spirit and scope of the technical solutions of the embodiments of this application, and should all be included within the protection scope of this application.

Claims

1. A method for optimizing a principal curve based on feature constraints, characterized in that, Including: Obtain high-dimensional non-linear data to be analyzed, and map the high-dimensional non-linear data to a target Riemannian manifold; The high-dimensional non-linear data has specific attributes; Construct an initial principal curve to be optimized through data points of the target Riemannian manifold; Determine a target projection point where a target data point is projected to a target position; The target data point is any data point of the target Riemannian manifold, and the target position is any position of the principal curve to be optimized; Determine a joint contribution factor of the target data point according to a spatial constraint coefficient and at least one data attribute feature constraint coefficient; wherein, the spatial constraint coefficient is determined based on the Riemannian distance between the target projection point and a target curve point, the data attribute feature constraint coefficient is determined based on the specific attribute, and the target curve point is the point on the principal curve to be optimized at the target position; Update the principal curve to be optimized by minimizing a target optimization error to obtain a final target principal curve; wherein, the target optimization error is weighted by the joint contribution factor.

2. The method for optimizing a principal curve based on feature constraints according to claim 1, wherein The specific attributes include a time series attribute and a distribution attribute, the time series attribute at least corresponds to a time feature, and the distribution attribute at least corresponds to a density feature.

3. The main curve optimization method based on feature constraints according to claim 1, wherein When the high-dimensional non-linear data is human motion data, the step of obtaining the high-dimensional non-linear data to be analyzed and mapping the high-dimensional non-linear data to a target Riemannian manifold includes: Obtain joint point information and joint limb information of a human skeleton; Obtain rotation data of two adjacent joint points at the same moment from the human motion data, and convert the rotation data into a rotation matrix of a joint limb; Determine all rotation matrices at the same moment as a group of rotation matrix sequences; wherein, multiple groups of rotation matrix sequences at different moments are sequential in time; Map the multiple groups of rotation matrix sequences to the target Riemannian manifold.

4. The main curve optimization method based on feature constraints according to claim 1, wherein The step of constructing a principal curve to be optimized through data points of the target Riemannian manifold includes: Divide data points on the Riemannian manifold into multiple data groups; Respectively determine mean points of the data groups based on Riemannian geometry; Generate intermediate points by geodesic interpolation between two adjacent mean points; Fit the principal curve to be optimized according to the mean points and the intermediate points.

5. The method for optimizing the main curve based on feature constraints according to claim 4, wherein When the data points correspond to rotation matrices, the step of respectively determining mean points of the data groups based on Riemannian geometry includes: Use logarithmic mapping to map each rotation matrix in the data group to a rotation vector in Lie algebra; Calculate an average vector based on the rotation vectors; Use exponential mapping to map the average vector back to Lie group to obtain an initial candidate mean point; Calculate the Riemannian distance from each rotation matrix in the data group to the candidate mean point; The candidate mean points are updated by minimizing the objective function to obtain the mean points of the final data grouping, and the objective function to be minimized is: ; Among them, represents the mean point, and R represents the candidate mean point, represents the rotation matrix in the data grouping, and SO(3) represents the three-dimensional rotation group, represents the Riemannian distance from the rotation matrix to the candidate mean point.

6. The method for optimizing a principal curve based on feature constraints according to claim 1, wherein Determine the joint contribution factor of the target data point through the following formula: ; where t represents the position parameter of the main curve, and n represents the identification parameter of the data points, represents the combined contribution factor, represents the spatial constraint coefficient, represents the data attribute feature constraint coefficient, represents the smoothing kernel function.

7. The method for optimizing a principal curve based on feature constraints according to claim 6, wherein When the specific attribute is a timing attribute, the combined contribution factor of the target data point is determined by the following formula: ; Among them, represents the time feature constraint coefficient, represents the curve point at parameter t on the main curve, represents the data point the projection point of the data point at parameter t on the main curve, represents the Riemannian distance between the projection point and the curve point, represents the spatial scale factor, represents the time scale factor, represents the time deviation between the data point and the curve point, represents the time corresponding to the curve point with parameter t on the main curve, represents the data point the corresponding time.

8. The method for optimizing the master curve based on feature constraints according to claim 1, wherein When the high-dimensional non-linear data contains rotation information, the principal curve optimization method based on feature constraint further includes determining an adaptive rotation axis direction of the high-dimensional non-linear data, and the step of determining the adaptive rotation axis direction of the high-dimensional non-linear data includes: Convert the high-dimensional non-linear data into at least one rotation matrix dataset; the rotation matrix dataset includes a plurality of rotation matrices; For any one rotation matrix dataset, determine the mean matrix of the rotation matrix dataset based on Riemannian geometry; Obtain the deviation vectors of the mean matrix and each rotation matrix in the rotation matrix dataset; Construct a covariance matrix based on the deviation vectors; Perform eigenvalue decomposition on the covariance matrix to obtain eigenvalues and corresponding eigenvectors; the eigenvalues represent the amount of change in the direction of the corresponding eigenvectors, and the eigenvectors are used to represent the change direction of the rotation matrix dataset; Select the eigenvector corresponding to the smallest eigenvalue and determine it as the rotation axis direction of the rotation matrix dataset; Determine the adaptive rotation axis direction of the high-dimensional non-linear data based on the rotation axis directions of all different rotation matrix datasets.

9. An apparatus for optimizing a principal curve based on feature constraints, characterized in that Comprising: A data acquisition module, configured to acquire high-dimensional non-linear data to be analyzed and map the high-dimensional non-linear data to a target Riemannian manifold; The high-dimensional non-linear data has specific attributes; A curve construction module, configured to construct an initial main curve to be optimized through the data points of the target Riemannian manifold; A projection module, configured to determine a target projection point where a target data point is projected to a target position; The target data point is any data point on the target Riemannian manifold, and the target position is any position on the main curve to be optimized; A contribution calculation module, configured to determine a joint contribution factor of the target data point according to a spatial constraint coefficient and at least one data attribute feature constraint coefficient; wherein, the spatial constraint coefficient is determined based on the Riemannian distance between the target projection point and the target curve point, the data attribute feature constraint coefficient is determined based on the specific attribute, and the target curve point is the point on the main curve to be optimized at the target position; A curve update module, configured to update the main curve to be optimized by minimizing a target optimization error to obtain a final target main curve; wherein, the target optimization error is weighted by the joint contribution factor.

10. An electronic device, comprising a memory, a processor, and a computer program stored in the memory and executable on the processor, characterized in that, When the processor executes the computer program, it implements the feature constraint-based main curve optimization method according to any one of claims 1 to 8.

Citation Information

Patent Citations

  • Riemannian manifold discretization geodesic calculation method based on deep learning

    CN117708577A

  • Scene elevation difference high-precision monitoring method based on Riemannian manifold

    CN119917775A

  • Isogeometric analysis parameterization migration method based on discrete geometric mapping

    WO2025065830A1