An online matching optimization method, device, medium and system combining geometry and texture

By combining online matching optimization methods of geometry and texture, and using depth texture images for pose estimation and optimization, the problem of poor adaptability of traditional 3D scanning systems to optical changes is solved, and more robust and accurate registration results are achieved.

CN115170634BActive Publication Date: 2026-02-03SHENZHEN JIMUYIDA TECH CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202210772666.3
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2021-06-04
Publication Date
2026-02-03
Estimated Expiration
2041-06-04

AI Technical Summary

Technical Problem

Traditional 3D scanning methods are poorly adaptable to optical changes, highly susceptible to environmental interference, have unstable registration results, and large pose estimation errors.

Method used

An online matching optimization method combining geometry and texture is used to acquire depth texture images through a 3D scanning device. Preliminary pose estimation and optimization are performed using depth geometric information and texture information. Iterative calculations are performed using a nonlinear optimization method, and geometric and optical constraints are integrated to gradually refine the pose estimation and improve its accuracy.

Benefits of technology

This improves the adaptability of the 3D scanning system to optical changes, enhances its anti-interference ability, and ensures the robustness and accuracy of the registration results.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN115170634B_ABST
    Figure CN115170634B_ABST
Patent Text Reader

Abstract

The application provides an online matching optimization method combining geometry and texture and a three-dimensional scanning system, and the method comprises the following steps: acquiring a depth texture image of a target object by using a three-dimensional scanning device; obtaining a preliminary pose estimation of the three-dimensional scanning device according to the depth texture image information; and performing optimization on the preliminary pose based on the depth geometry information and the texture information of the image to obtain a refined inter-frame motion estimation. The method fuses geometric and optical double constraints, fully utilizes texture information, proposes to calculate and solve texture images, obtains feature values which are not sensitive to light and have strong anti-interference ability to replace unprocessed pixel intensity, so that the adaptability of the system to optical changes is stronger, the registration result is more robust, and the inter-frame matching efficiency and accuracy are improved.
Need to check novelty before this filing date? Find Prior Art

Description

[0001] This application is a divisional application of application number 202110625611.5, the parent application which was filed on June 4, 2021, and is entitled "An Online Matching Optimization Method and a 3D Scanning System Combining Geometry and Texture". Technical Field

[0002] This invention belongs to the field of image recognition, and more specifically, relates to an online matching optimization method, device, medium, and system that combines geometry and texture. Background Technology

[0003] In recent years, 3D scanning, as a rapid 3D digitization technology, has been increasingly used in various fields, including reverse engineering, industrial inspection, computer vision, CG production, etc. Especially in the rapidly developing fields of 3D printing and intelligent manufacturing, 3D scanning, as a front-end 3D digitization and 3D vision sensing technology, has become an important link in the industrial chain. At the same time, various applications have put forward higher requirements for 3D scanning in terms of cost, practicality, accuracy and reliability.

[0004] Traditional 3D scanning methods, such as the direct method, use pixel grayscale for matching, which results in poor adaptability to optical changes, significant environmental interference, unstable registration results, and large pose estimation errors. Summary of the Invention

[0005] This invention provides an online matching optimization method and a 3D scanning system that combines geometry and texture to improve scanning efficiency.

[0006] The technical solution adopted by this invention to solve its technical problem is: an online matching optimization method combining geometry and texture, the method comprising:

[0007] A 3D scanning device is used to acquire depth and texture images of the target object;

[0008] Based on the depth texture image information, a preliminary pose estimate of the three-dimensional scanning device is obtained;

[0009] The preliminary pose is optimized based on the depth geometric information and texture information of the image to obtain a refined inter-frame motion estimate;

[0010] Image matching optimization is performed based on the inter-frame motion estimation.

[0011] Acquiring depth and texture images of a target object using a 3D scanning device includes: acquiring depth information and texture information of the target object through an alternating projection method using the depth acquisition device and the camera device.

[0012] The step of obtaining a preliminary pose estimate of the scanning device based on the depth texture image information includes:

[0013] For each image frame that needs to be matched in the depth texture image, a sample frame that is compatible with the image frame is obtained; for each image frame and sample frame, the corresponding image feature data is extracted, and image feature matching is performed between the image frame and the corresponding sample frame to obtain multiple initial feature pairs; an initial transformation matrix is ​​selected from the multiple initial feature pairs, and the initial pose of the three-dimensional scanning device is estimated based on the initial transformation matrix.

[0014] The preliminary pose is optimized based on the depth geometric information and texture information of the image to obtain the inter-frame motion estimation of the image, including:

[0015] Using the initial transformation matrix as the optimization objective, an initial optimization function is constructed; a nonlinear optimization method is used to iteratively optimize the optimization objective to obtain the refined inter-frame motion estimate.

[0016] The initial optimization, using the initial transformation matrix as the optimization objective, includes:

[0017] An initial optimization function is constructed based on geometric and optical constraints; iterative optimization is then performed based on the gradient information of the current frame and sample frames.

[0018] This invention also discloses a three-dimensional scanning device, comprising:

[0019] The acquisition unit is used to acquire the depth texture image of the target object scanned by the 3D scanning device;

[0020] An inter-frame motion estimation module is used to obtain a preliminary pose estimate of the 3D scanning device based on the depth texture image information; and,

[0021] The preliminary pose is optimized based on the depth geometric information and texture information of the image to obtain a refined inter-frame motion estimate;

[0022] The matching optimization module is used to perform image matching optimization based on the inter-frame motion estimation.

[0023] The present invention also discloses a computer-readable storage medium, characterized in that it stores a computer program, which, when executed by a processor, implements the steps of the above-described method.

[0024] The present invention also discloses a three-dimensional scanning system, including a memory and a processor, wherein the memory stores a computer program, and the processor executes the computer program to implement the steps of the above-described method.

[0025] The online matching optimization method, apparatus, medium, and system combining geometry and texture of this invention integrates both geometric and optical constraints, fully utilizes texture information, and proposes to calculate and solve texture images to obtain feature values ​​that are insensitive to illumination and have strong anti-interference capabilities, replacing unprocessed pixel intensities. This makes the system more adaptable to optical changes and the registration results more robust. Furthermore, it employs a stepwise refinement strategy to decompose and simplify the complex problem. First, it preliminarily estimates the pose using features, then refines the pose to gradually obtain an accurate pose estimate, achieving rapid and accurate optimization. Attached Figure Description

[0026] The present invention will be further described below with reference to the accompanying drawings and embodiments. In the accompanying drawings:

[0027] Figure 1 This is a flowchart of an online matching optimization method combining geometry and texture, according to one embodiment of the present invention.

[0028] Figure 2 This is a typical optical path diagram of a 3D scanning system;

[0029] Figure 3 This is a schematic diagram showing the process details of online matching optimization combining geometry and texture in one embodiment of the present invention;

[0030] Figure 4 This is a schematic diagram of the effect after the network is integrated. Detailed Implementation

[0031] To provide a clearer understanding of the technical features, objectives, and effects of the present invention, specific embodiments of the present invention will now be described in detail with reference to the accompanying drawings.

[0032] Please refer to Figure 1 This is a flowchart of an online matching optimization method combining geometry and texture in one embodiment of the present invention. The method includes:

[0033] S1. Use a 3D scanning device to acquire the depth texture image of the target object;

[0034] Specifically, a one-to-one corresponding depth-texture image pair is obtained, the depth-texture image pair including a depth image acquired by a depth sensor and a texture image acquired by a camera device.

[0035] Please refer to Figure 2This diagram illustrates a typical optical path of a 3D scanning system. Two optical paths exist: Beam A is a structured light, which, after penetrating a specific coded pattern as white light, is projected onto the object being measured. Beam B is a texture illumination light, which, also as white light, is directly projected onto the object. Simultaneously with beam B's projection, the camera activates its image capture function, with its exposure time strictly synchronized with the beam projection time pulse. It should be noted that while beam A is being projected once, the camera also captures a single image of the object projected by beam A. Immediately afterward, beam B is projected, and the camera captures a single image of the object projected by beam B. This constitutes a single cycle of the measurement process. By repeatedly performing this process at a certain repetition frequency, while the relative positions and angles of the 3D scanning device and the object being measured continuously change, continuous measurement of the object's 3D scanning structure can be achieved.

[0036] Optionally, in one embodiment, the aforementioned three-dimensional scanning device is applied in a continuous rapid measurement mode. In this mode, beams A and B are projected alternately to complete the measurement of the object being measured. The beams emitted by the three-dimensional scanning device are output in the form of high-power short pulses, which provides a good foundation for subsequent high-precision measurements. It should be noted that in this embodiment, the instantaneous power of beam A can reach the kilowatt level, with a pulse width in the hundreds of microseconds range; the instantaneous power of beam B is in the hundreds of watts range, with a pulse width in the hundreds of microseconds range; the time difference between beams A and B and the camera exposure time for both are in the hundreds of microseconds range.

[0037] S2. Based on the depth texture image information, obtain the preliminary pose estimate of the three-dimensional scanning device;

[0038] Specifically, a stepwise refinement strategy is adopted to perform feature matching between the depth texture image corresponding to the current frame and the depth texture image corresponding to the sample frame in order to estimate the preliminary pose of the depth sensor in the 3D scanning system.

[0039] Furthermore, the estimation of the preliminary pose of the depth sensor includes:

[0040] S21. For each image frame that needs to be matched in the depth texture image, obtain a sample frame that is compatible with the image frame.

[0041] S22. For each of the image frames and sample frames, extract the corresponding image feature data, and perform image feature matching between the image frames and the corresponding sample frames to obtain multiple initial feature pairs;

[0042] S23. Select an initial transformation matrix from the plurality of initial feature pairs, and estimate the preliminary pose of the depth sensor based on the initial transformation matrix.

[0043] Specifically, this application considers extracting SIFT features from captured RGB images and performing feature matching between the current frame and sample frames based on the extracted SIFT features. It should be noted that SIFT is a widely used feature detector and descriptor, significantly superior to other features in terms of the detail and stability of feature point description. During the SIFT matching process, features are extracted from image frames F... i Find the nearest neighbor to obtain frame F j The best candidate match for each keypoint. This brute-force matching method can obtain frame F. j With frame F i The initial N feature pairs between the two are represented by vectors (U;V). These feature pairs contain both correct data (Inliers) and outliers. In order to filter out the correct data from these matched feature pairs, this application uses the RANSAC algorithm to filter the effective sample data from the sample dataset containing outliers. The idea of ​​the RANSAC algorithm is: randomly select a set of RANSAC samples from N and calculate the transformation matrix (r;t). Based on (r;t), calculate the number of consistent points that satisfy the preset error metric function (see formula (1) below), that is, the number of inliers f, as shown in formula (2) below. This process is repeated cyclically to obtain the consistent set with the largest f. Then, the optimal transformation matrix is ​​calculated from the consistent set.

[0044]

[0045]

[0046] Among them, I(U i V i (,r,t) represents the i-th matching point pair (U i V i If the preset threshold values ​​d and θ can be satisfied under the current constraints (r; t), then I = 1; otherwise, I = 0. Pi N Qi Representing three-dimensional point P respectively i Q i The unit normal vector. N is the total number of matched point pairs. f(r,t) is the number of interior points.

[0047] S3. Based on the depth geometric information and texture information of the image, the preliminary pose is optimized to obtain a refined inter-frame motion estimate.

[0048] Specifically, the preliminary pose estimated in step S2 is optimized by combining geometric constraints and texture constraints to obtain a refined inter-frame motion estimate.

[0049] The optimization of the preliminary pose estimated in step S2 by combining geometric and texture constraints to obtain a refined inter-frame motion estimate includes:

[0050] S31. Using the initial transformation matrix as the optimization objective, construct the initial optimization function E1 according to the following formula:

[0051]

[0052] Where G represents geometric constraints, L represents texture constraints, ω represents the confidence level of texture constraints, and κ represents the geometric constraints. i,j Let p be a set of matching point pairs, p be a 3D point in image frame i, q be the corresponding point of 3D point p in image frame j, and m be the preset total number of matching point pairs.

[0053] Specifically, combining geometric and optical constraints, the minimization objective in the current embodiment includes two parts: one is the distance between the tangent plane of each target point and its corresponding source point, and the other is the gradient error between each target point and its corresponding source point. The two will be assigned different weights w according to the actual application.

[0054] In one embodiment, iterative optimization is performed based on gradient information between the current frame and sample frames.

[0055] Specifically, in the current frame F i With sample frame F j The corresponding matching point set κ i,j In the above, assume that p = (p x ,p y ,p z ,1)T is the source point cloud, q=(p x ,p y ,p z ,1) T For the target point cloud corresponding to p, n = (n x ,n y ,n z 1) T is the unit normal vector, g p Let g be the gradient value of the source point cloud p. q Let q be the gradient value of the target point cloud, and m be the number of matching point pairs. When iteratively optimizing the above formula (3), the goal of each iteration is to find the optimal (r) opt ;t opt ), where (r opt ;t opt The following equation must be satisfied:

[0056]

[0057] S32. The optimization objective is iteratively optimized using a nonlinear optimization method, and when the preset iteration termination condition is reached, the refined inter-frame motion estimate is obtained based on the optimal transformation matrix output by the last iteration.

[0058] Specifically, in order to solve the objective function constructed above, this embodiment defines the initial transformation matrix as a vector with six parameters: that is, ξ = (α, β, γ, a, b, c). Then the initial transformation matrix can be linearly represented as:

[0059]

[0060] Among them, T k This is the transform estimate from the last iteration, currently using the Gauss-Newton method. r T J r ξ=-J r T r Solve for the parameter ξ and apply the parameter ξ to T. k To update T, where r is the residual and J r It is a Jacobian matrix.

[0061] In one embodiment, the preset iteration termination condition can be reaching a preset maximum number of iterations, etc., and different embodiments can be flexibly adjusted according to the actual application scenario.

[0062] In the above embodiments, geometric and optical constraints are integrated, texture information is fully utilized, and a method is proposed to calculate and solve the texture image to obtain feature values ​​that are insensitive to illumination and have strong anti-interference ability to replace the unprocessed pixel intensity. This makes the system more adaptable to optical changes and the registration results more robust.

[0063] Image matching optimization based on the inter-frame motion estimation includes:

[0064] S4. The data obtained through inter-frame motion estimation is segmented to obtain multiple data segments, and the pose in each data segment is optimized; wherein each data segment includes multiple image frames.

[0065] S5. For each data segment, select a key frame from the multiple image frames included in the data segment, and combine the key frames and loop closure information to perform joint optimization between segments.

[0066] Optionally, keyframe extraction needs to satisfy at least one of the following conditions:

[0067] (1) There is at least one keyframe in every N image frames, so that global information can be expressed through the keyframe.

[0068] (2) When the current image frame can match the previous image frame, but the current image frame cannot match the preset reference key frame, the previous image frame of the current image frame will be added to the preset key frame set to ensure the continuity of trajectory tracking.

[0069] (3) Although the current image frame can match the previous image frame and the current image frame can also match the preset reference key frame, the overlap rate between the current image frame and the preset reference key frame is not high enough. In this case, the current image frame needs to be added to the preset key frame set to ensure that there is overlap between adjacent key frames.

[0070] In one embodiment, the absolute pose estimation of the depth sensor accumulates significant pose errors over time. Furthermore, after implementing the local optimization measures in step S4, the pose information between segments lacks global consistency, and accumulated errors persist. To overcome these problems, this embodiment utilizes loop closure information and each keyframe to perform joint optimization between segments. It should be noted that loop closure information is typically calculated directly based on the image or features. In one embodiment, to obtain accurate loop closure information, this application employs an inter-frame matching method, matching adjacent keyframes pairwise. A corresponding loop is formed when a match is successful and the overlap rate reaches a set threshold.

[0071] In addition, since keyframes run through the entire tracking process and can fully reflect the global picture, in order to improve the efficiency of pose optimization, the global optimization in this embodiment does not involve all frames. Instead, one frame is selected from each data segment to represent that data segment. This image frame is collectively referred to as the keyframe. Then, the loop closure information is combined to perform global optimization. At this time, most of the accumulated errors can be quickly eliminated through global optimization.

[0072] In the above embodiments, based on the segmented multi-mode optimization strategy, the problem can be modeled at different levels of abstraction, achieving fast and accurate optimization.

[0073] S6. For each data segment, fix the pose of the key frame in the corresponding data segment, and optimize the pose of other image frames in the data segment to obtain a globally consistent and smooth motion trajectory map.

[0074] Specifically, the poses of keyframes have been updated through optimization of the global pose graph. However, to obtain a globally consistent and smooth motion trajectory, the poses within a local area also need to be updated. Therefore, this embodiment adopts a layered approach, not optimizing all image frames simultaneously, but fixing the poses of each keyframe segment and only optimizing the poses of other image frames within the segment.

[0075] S7. Combining the relative pose measured by the depth sensor and the absolute pose estimated by the motion trajectory map, construct a corresponding target optimization function.

[0076] S8. Incorporate the preset penalty factor into the target optimization function, and through iterative transformation estimation, eliminate the cumulative error generated as the number of scan frames increases during inter-frame matching, and perform fusion mesh construction.

[0077] Specifically, in step S8, when incorporating the preset penalty factor into the objective optimization function, the expression formula of the objective optimization function E2 is as follows:

[0078] E2=∑ i,j ρ(e 2 (p i ,p j ;∑ i,j ,T i,j (6)

[0079] Here, the estimated absolute pose is taken as a node, p i Representing nodes i and p j Represents node j; T i,j Represents the relative pose between node i and node j, ∑ i,j This indicates summing over all constraint pairs; e 2 (p i ,p j ;∑ i,j ,T i,j )=e(p i ,p j ;T i,j ) T ∑ i,j -1 e(p i ,p j ;T i,j ), e(p i ,p j ;T i,j ) = T i,j -p i -1 p j ρ is the penalty factor for integration.

[0080] In one embodiment, u = d 2 d represents the surface diameter of the reconstructed object. Considering that a suitable penalty function can effectively perform verification and filtering without increasing computational cost, the current embodiment uses the Geman-mclure function from M-estimation, i.e.

[0081] Since the above formula (6) is difficult to optimize directly, we currently assume relation l, and assume the objective optimization function E2 is:

[0082] E2=∑ i,j l(e 2 (p i ,p j ;∑ i,j ,T i,j )+∑ i,j ψ(l); (7)

[0083] Among them, it is known Minimize the formula E2 by taking the partial derivative with respect to l. In actual calculations, l is considered as the confidence level. Constraints with smaller residuals have higher error weights and are considered more reliable; conversely, constraints with larger residuals are less reliable. This serves as a verification and elimination mechanism, resulting in robust optimization. Furthermore, the selection of the parameter μ is crucial; μ = d 2 , representing the surface diameter of the reconstructed object, controls the range of significant influence of the residuals on the objective. A larger μ makes the objective function smoother and allows more corresponding terms to participate in the optimization. As μ decreases, the objective function becomes sharper, more outlier matches are eliminated, and the data participating in the optimization becomes more accurate.

[0084] To solve this nonlinear squared error function problem, the transformation matrix is ​​also transformed according to formula (5) in the current embodiment. Considering that only a small number of nodes in the pose graph have direct edge connections, i.e., the sparsity of the pose graph, and for numerical stability, the current embodiment uses a sparse BA algorithm to solve it. Sparse BA is usually optimized using the LM method. LM adds a positive definite diagonal matrix to the Gauss-Newton method, i.e., through (J r T J r +λI)ξ=-J r T r Let's solve for ξ.

[0085] It should be noted that the effect after rapid optimization is as follows: Figure 4 As shown in (c), further fusion and mesh construction are performed, and the effect is as follows. Figure 4 As shown in (d).

[0086] In one embodiment, a 3D scanning system applied to the online matching optimization method is also provided, the system comprising:

[0087] The acquisition module is used to acquire a one-to-one corresponding depth texture image pair, wherein the depth texture image pair includes a depth image acquired by a depth sensor and a texture image acquired by a camera device.

[0088] The inter-frame motion estimation module is used to perform feature matching between the depth texture image pair corresponding to the current frame and the depth texture image pair corresponding to the sample frame using a progressive refinement strategy to estimate the initial pose of the depth sensor, and to optimize the estimated initial pose by combining geometric constraints and texture constraints to obtain a refined inter-frame motion estimate.

[0089] A multi-mode optimization module is used to segment the data obtained through inter-frame motion estimation to obtain multiple data segments, and optimize the pose of each data segment; wherein each data segment includes multiple image frames; for each data segment, a key frame is selected from the multiple image frames included in the data segment, and combined with the key frame and loop closure information, joint optimization between segments is performed, and the pose of the key frame in the corresponding data segment is fixed, and the pose of other image frames in the data segment is optimized to obtain a globally consistent and smooth motion trajectory map;

[0090] The cumulative error elimination module is used to construct a corresponding target optimization function by combining the relative pose measured by the depth sensor and the absolute pose estimated by the motion trajectory map; it is also used to incorporate a preset penalty factor into the target optimization function, and eliminate the cumulative error generated as the number of scan frames increases during inter-frame matching through iterative transformation estimation, and perform fusion mesh construction.

[0091] In one embodiment, the depth sensor includes a projection module and a depth information acquisition module. The projection module is used to project white light or a structured light beam of a specific wavelength onto the surface of the object being measured. The depth information acquisition module is used to acquire depth information of the surface of the object being measured when the projection module projects the structured light beam. The camera device includes a texture information acquisition module, which is used to acquire texture information of the surface of the object being measured when the camera device projects a texture illumination beam onto the surface of the object being measured.

[0092] Here, the structured beam and the textured illumination beam are projected alternately. When the structured beam is projected once, the depth information acquisition module acquires the depth information of the surface of the object being measured. Then, the projection of the textured illumination beam is started, and the camera device acquires the texture information of the surface of the object being measured projected by the textured illumination beam once.

[0093] The above is a single cycle of the measurement process of texture and depth information of the surface of the object being measured. When the above measurement process is repeated at a certain repetition frequency, the relative position and relative angle between the camera device, the depth sensor and the object being measured will change continuously, thus completing the continuous measurement of the structure of the object being measured.

[0094] In one embodiment, a computer-readable storage medium is also provided, on which a computer program is stored, which, when executed by a processor, implements the steps of any of the above-described online matching optimization methods.

[0095] In one embodiment, a 3D scanning device for an online matching optimization method is also provided, including a memory and a processor, wherein the memory stores a computer program, and the processor executes the computer program to implement the steps in the above-described method embodiments.

[0096] This invention discloses an online matching optimization method and 3D scanning system that combines geometry and texture. On one hand, it integrates both geometric and optical constraints, fully utilizing texture information. It proposes to calculate and solve for texture images, obtaining feature values ​​that are insensitive to illumination and have strong anti-interference capabilities to replace unprocessed pixel intensities. This makes the system more adaptable to optical changes and the registration results more robust. On the other hand, it employs a stepwise refinement strategy to decompose and simplify the complex problem. First, it preliminarily estimates the pose using features, then refines the pose to gradually obtain an accurate pose estimate. Furthermore, a penalty factor is added to the subsequently established optimization objective function. Without increasing additional computational costs, it can effectively verify and filter different constraint pairs, ensuring the accuracy and stability of the optimization. Finally, it employs a piecewise multi-mode optimization strategy, enabling modeling of the problem at different levels of abstraction, achieving fast and accurate optimization.

[0097] The embodiments of the present invention have been described above with reference to the accompanying drawings. However, the present invention is not limited to the specific embodiments described above. The specific embodiments described above are merely illustrative and not restrictive. Those skilled in the art can make many other forms under the guidance of the present invention without departing from the spirit and scope of the claims. All of these forms are within the protection scope of the present invention.

Claims

1. An online matching optimization method combining geometry and texture, characterized in that, The method includes: A 3D scanning device is used to acquire depth and texture images of the target object; Based on the depth texture image information, a preliminary pose estimate of the three-dimensional scanning device is obtained; The preliminary pose is optimized based on the depth geometric information and texture information of the image to obtain a refined inter-frame motion estimate. Image matching optimization is performed based on the inter-frame motion estimation; The step of obtaining a preliminary pose estimate of the scanning device based on the depth texture image information includes: For each image frame that needs to be matched in the depth texture image, obtain sample frames that are compatible with the image frames; For each of the image frames and sample frames, corresponding image feature data is extracted, and image feature matching is performed between the image frames and the corresponding sample frames to obtain multiple initial feature pairs; An initial transformation matrix is ​​selected from the plurality of initial feature pairs, and the initial pose of the three-dimensional scanning device is estimated based on the initial transformation matrix; The step of optimizing the preliminary pose based on the depth geometric information and texture information of the image to obtain a refined inter-frame motion estimate includes: Using the initial transformation matrix as the optimization objective, an initial optimization function is constructed; The optimization objective is iteratively optimized using a nonlinear optimization method to obtain a refined inter-frame motion estimate. The step of constructing an initial optimization function with the initial transformation matrix as the optimization objective includes: An initial optimization function is constructed by combining geometric and optical constraints; Iterative optimization is performed based on the gradient information of the current frame and the sample frames; The iterative formula is as follows: Among them, (r opt ;t opt ) represents the optimal transformation matrix after iteration; p is a 3D point in image frame i, q is the corresponding point of 3D point p in image frame j; m is the number of matching point pairs; ω is the confidence score of the texture constraint; k ij For matching point pairs; g p Let g be the gradient value of the source point cloud p. q Let q be the gradient value of the target point cloud, and (r; t) be the transformation matrix.

2. The method according to claim 1, characterized in that, The three-dimensional scanning device includes a depth acquisition device and a camera device. The process of acquiring a depth texture image of the target object using the three-dimensional scanning device includes: The depth information and texture information of the target object are acquired by alternating projection of the depth acquisition device and the camera device.

3. A three-dimensional scanning device for applying the method according to any one of claims 1 to 2, said three-dimensional scanning device comprising: The acquisition unit is used to acquire the depth texture image of the target object scanned by the 3D scanning device; The inter-frame motion estimation module is used to obtain a preliminary pose estimate of the three-dimensional scanning device based on the depth texture image information. as well as, The preliminary pose is optimized based on the depth geometric information and texture information of the image to obtain a refined inter-frame motion estimate; The matching optimization module is used to perform image matching optimization based on the inter-frame motion estimation.

4. A computer-readable storage medium, characterized in that, The device contains a computer program that, when executed by a processor, implements the steps of the method described in any one of claims 1 to 2.

5. A three-dimensional scanning system, comprising a memory and a processor, wherein the memory stores a computer program, characterized in that, When the processor executes the computer program, it implements the steps of the method according to any one of claims 1 to 2.

Citation Information

Patent Citations

  • Texture mapping method terminal and device based on gradient domain, and medium

    CN111553969A

  • Fast iterative registration method based on initial value, medium, terminal and device

    CN111739071A