Image-based ophthalmic robot control method, system, equipment and medium

By combining dynamic registration of three-dimensional OCT images and visual evoked potential signals with neural network prediction, the problem of insufficient intraoperative optic nerve function monitoring was solved, individualized prediction of postoperative vision recovery and real-time optimization of surgical operations were achieved, and the safety and effectiveness of the surgery were improved.

CN120770936AActive Publication Date: 2025-10-14SHANGHAI LOHAS YUAN MEDICAL TECHNOLOGY CO LTD

Patent Information

Application Number
CN202511286878.0
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-09-10
Publication Date
2025-10-14
Estimated Expiration
2045-09-10

AI Technical Summary

Technical Problem

Existing technologies are unable to capture the electrophysiological signals of the optic nerve in real time, making it difficult to optimize surgical strategies in real time based on individualized functional recovery needs. There is also a lack of dynamic monitoring of the risk of intraoperative visual function damage and prediction of postoperative vision recovery effects.

Method used

By acquiring intraoperative three-dimensional OCT image sequences and visual evoked potential signals, a spatial model of the optic nerve is generated, and dynamic alignment is performed through an eye movement compensation algorithm. Combined with a pre-trained neural network model, the risk coefficient of visual function damage and the predicted value of postoperative vision recovery are calculated, and robot operation correction control instructions are dynamically generated.

Benefits of technology

It realizes dynamic monitoring and quantitative correlation prediction of intraoperative optic nerve function, improves the individualized precision of surgical operation and the effect of visual function protection, and ensures surgical safety and postoperative vision recovery.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120770936A_ABST
    Figure CN120770936A_ABST
Patent Text Reader

Abstract

The invention provides an image-based ophthalmology robot control method, system and device and a medium, and the method comprises the steps: obtaining an intraoperative three-dimensional OCT image sequence and a visual evoked potential signal, and segmenting an optic nerve region to generate an optic nerve space model; performing dynamic registration on the optic nerve space model and the visual evoked potential signal according to an eyeball motion compensation algorithm to generate a fusion data volume; calculating a visual function injury risk coefficient and a postoperative vision recovery predicted value through a pre-trained neural network prediction model based on the fusion data body and the real-time operation parameters of the robot end instrument; and dynamically generating a robot operation correction control instruction according to the injury risk coefficient and the postoperative vision recovery predicted value. By adopting the method, the individualized precision of an operation scheme and the postoperative visual function protection effect can be improved.
Need to check novelty before this filing date? Find Prior Art

Description

TECHNICAL FIELD

[0001] The application belongs to the technical field of ophthalmic robot control, and particularly relates to an image-based ophthalmic robot control method, system, device and medium. BACKGROUND

[0002] With the development of ophthalmic surgical robots and medical imaging technology, control systems guided by images such as OCT and intraoperative microscopes can achieve precise positioning of surgical instruments in intraocular tissues. By segmenting key structures such as the retina and optic nerve, and combining robot motion planning algorithms, the operation safety is improved. In traditional methods, postoperative clinical follow-up is completely relied on for surgical outcome evaluation, and the surgeon only adjusts the instrument parameters according to the anatomical position and operating experience during the operation, lacking real-time prediction ability of the postoperative visual function recovery effect of the patient.

[0003] However, damage to the optic nerve function is often caused by irreversible intraoperative operations (such as excessive compression and energy accumulation), and the existing technology cannot dynamically capture the optic nerve electrophysiological signals (such as visual evoked potentials), nor has it established a predictive correlation between intraoperative operation parameters and postoperative vision, making it difficult to optimize the surgical strategy in real time according to individual functional recovery needs. SUMMARY

[0004] Therefore, it is necessary to provide an image-based ophthalmic robot control method, system, device and medium, which can construct a predictive correlation between operation parameters and postoperative visual recovery effect through real-time fusion analysis of intraoperative optic nerve bioelectric signals and multi-modal images, to improve the individual accuracy of surgical plans and the postoperative visual function protection effect.

[0005] In a first aspect, the present application provides an image-based ophthalmic robot control method, comprising: obtaining an intraoperative three-dimensional OCT image sequence and a visual evoked potential signal, segmenting the optic nerve region to generate an optic nerve spatial model; dynamically registering the optic nerve spatial model and the visual evoked potential signal according to an eye movement compensation algorithm to generate a fusion data body; based on the fusion data body and real-time operation parameters of a robot end instrument, calculating a visual function damage risk coefficient and a postoperative visual recovery prediction value through a pre-trained neural network prediction model; generating a robot operation correction control instruction dynamically according to the damage risk coefficient and the postoperative visual recovery prediction value.

[0006] In one embodiment, dynamically registering the optic nerve spatial model and the visual evoked potential signal according to the eye movement compensation algorithm to generate the fusion data body comprises: performing optical flow field analysis on the three-dimensional OCT image sequence of adjacent frames to generate a tissue displacement vector field; constructing a dynamic correction matrix according to the tissue displacement vector field; mapping the coordinates of the visual evoked potential signal collection points to corresponding nodes of the optic nerve spatial model through the dynamic correction matrix; binding the amplitude parameter and the latency offset parameter of the visual evoked potential signal to the mapped nodes to generate a fusion data body.

[0007] In one embodiment, based on the fusion data body and the real-time operation parameters of the robot end instrument, the visual function damage risk coefficient and the postoperative visual acuity recovery prediction value are calculated through a pre-trained neural network prediction model, including: calculating a dynamic optic nerve function index based on the amplitude parameter, the latency offset parameter, and the real-time pressure change rate of the robot end instrument; inputting the dynamic optic nerve function index and the real-time pressure into a pre-trained convolutional recurrent neural network to output a visual function damage risk coefficient; calculating a postoperative visual acuity recovery prediction value based on the functional relationship between the visual function damage risk coefficient and the preoperative baseline visual acuity.

[0008] In one embodiment, a robot operation correction control instruction is dynamically generated according to the damage risk coefficient and the postoperative visual acuity recovery prediction value, including: generating a current surgical stage identifier through a pre-set surgical stage division rule based on the real-time position coordinates of the robot end instrument in the three-dimensional OCT image sequence; determining whether the current surgical stage identifier belongs to a pre-set high-risk stage set; if the current surgical stage identifier belongs to the pre-set high-risk stage set, calculating a maximum allowable movement speed and a pressure safety threshold value according to the visual function damage risk coefficient; if the postoperative visual acuity recovery prediction value is less than a pre-set recovery threshold value, generating a motion pause instruction and triggering an alarm signal.

[0009] In one embodiment, the dynamic correction matrix is constructed according to the tissue displacement vector field, further including: affine transformation modeling of the tissue displacement vector field to generate a spatial compensation parameter matrix, the spatial compensation parameter matrix including translation, rotation, and scaling parameters; performing spatial transformation operation on the original coordinates of the visual evoked potential signal collection points using the spatial compensation parameter matrix to generate corrected spatial coordinates; screening an effective mapping node set according to the Euclidean distance between the corrected spatial coordinates and the nodes in the optic nerve spatial model; updating the spatial coordinate distribution of the fusion data body based on the effective mapping node set.

[0010] In one embodiment, the training method of the convolutional recurrent neural network includes: constructing a training data set, the training data set containing preoperative three-dimensional OCT image features, intraoperative instrument operation parameter sequences, and postoperative clinical visual function evaluation results; generating instrument operation parameter and visual function damage correlation simulation data through a generative adversarial network, the correlation simulation data being used to expand the training data set; extracting the time sequence dependent features of the instrument operation parameters using a long short-term memory module, and weighting and fusing the time sequence dependent features and the image features through an attention mechanism; optimizing the network weight parameters through a back propagation algorithm, so that the KL divergence of the visual function damage risk coefficient output by the convolutional recurrent neural network and the postoperative clinical evaluation result converges within a set range.

[0011] In one of the embodiments, the image-based ophthalmic robot control method further comprises: In the postoperative verification stage, the visual function damage risk coefficient sequence recorded in the historical surgery process and the postoperative visual recovery prediction value of the key operation point are obtained; performing Gaussian mixture model clustering analysis on the visual function damage risk coefficient sequence and the postoperative visual recovery prediction value to generate a high-risk operation mode feature set; based on the high-risk operation mode feature set and the anatomical structure features of the preoperative three-dimensional OCT image, generating a prognosis optimization score report through a random forest algorithm, the prognosis optimization score report including surgery path planning suggestions and risk avoidance strategies.

[0012] In a second aspect, the present application also provides an image-based ophthalmic robot control system, comprising: an image segmentation module for obtaining an intraoperative three-dimensional OCT image sequence and a visual evoked potential signal, segmenting the optic nerve region to generate an optic nerve spatial model; a data fusion module for dynamically registering the optic nerve spatial model and the visual evoked potential signal according to an eye movement compensation algorithm to generate a fusion data body; a visual function prediction analysis module for calculating a visual function damage risk coefficient and a postoperative visual recovery prediction value based on the fusion data body and real-time operation parameters of a robot end instrument through a pre-trained neural network prediction model; a control instruction generation module for dynamically generating robot operation correction control instructions according to the damage risk coefficient and the postoperative visual recovery prediction value.

[0013] In a third aspect, the present application also provides a computer device comprising a memory and a processor, the memory storing a computer program, and the processor implementing the above-mentioned image-based ophthalmic robot control method when executing the computer program.

[0014] In a fourth aspect, the present application also provides a computer readable storage medium, which stores a computer program, and the computer program is executed by a processor to implement the image-based ophthalmic robot control method.

[0015] The image-based ophthalmic robot control method, system, device and medium solve the problem of intraoperative optic nerve function dynamic monitoring, and realize quantitative correlation prediction of surgical operation parameters and postoperative visual recovery effect by establishing a real-time fusion analysis mechanism of anatomical structure and bioelectric signal, form a prognosis-oriented intraoperative closed-loop control, thereby improving the individualized precision of surgical operation and the protection effect of visual function. BRIEF DESCRIPTION OF DRAWINGS

[0016] In order to more clearly illustrate the technical solutions in the embodiments of the present application or the related art, the drawings needed to be used in the embodiments or the related art description will be briefly introduced. Obviously, the drawings in the following description are only some embodiments of the present application, and other drawings can be obtained by those skilled in the art without creative labor.

[0017] Figure 1 The flowchart of the image-based ophthalmic robot control method provided by the embodiments of the present application is shown in the figure. Figure 2 The structure diagram of the image-based ophthalmic robot control system provided by the embodiments of the present application is shown in the figure. DETAILED DESCRIPTION

[0018] In order to make the purpose, technical solutions and advantages of the present application more clear, the present application will be further described in detail below in combination with the drawings and embodiments. It should be understood that the specific embodiments described herein are only used to explain the present application, and are not used to limit the present application.

[0019] First, the terms involved in the embodiments of the present application are briefly introduced.

[0020] Intraoperative three-dimensional OCT image sequence refers to a three-dimensional image data set of eye tissue continuously collected by an optical coherence tomography (OCT) device during surgery. The image sequence presents the cross-sectional and longitudinal sectional morphology of structures such as the retina and the optic nerve in layers with a resolution of microns, and provides real-time dynamic change information of anatomical structures during surgery. Unlike static images, its time continuity supports tissue displacement tracking and provides a spatial reference for motion compensation algorithms. In robotic surgery in ophthalmology, the sequence is a basic data source for constructing a spatial model of the optic nerve, ensuring the accuracy of operations on fine anatomical structures.

[0021] Visual evoked potential signal (VEP) refers to a cortical bioelectric signal induced by visual stimulation, which reflects the functional integrity of the optic nerve pathway. Intraoperative VEP signal is collected by placing electrodes at specific points on the patient's scalp. The core parameters include amplitude (reflecting the degree of synchronous excitation of nerve cells) and latency (reflecting the speed of nerve conduction). In the technical solution, VEP signal is used as a real-time functional monitoring indicator to quantify the instantaneous impact of surgical operations on the function of the optic nerve, making up for the deficiency of pure anatomical images in assessing nerve activity.

[0022] Optic nerve spatial model refers to a digital three-dimensional geometric model of the optic nerve generated based on three-dimensional OCT image segmentation. The model extracts the boundary of the optic nerve and reconstructs its three-dimensional topological structure through image segmentation algorithms, and can include the coordinate spatial position information of key areas such as the optic disc and the nerve fiber layer. Its value lies in abstracting anatomical structures into computer-processable mathematical entities, providing a stable reference system for dynamic registration and supporting collision avoidance calculations in robot path planning.

[0023] According to the above explanations of the terms, the implementation environment of the image-based ophthalmic robot control method provided in the embodiments of the present application is described. Illustratively, the implementation environment includes a terminal, a sensor array, a processor, and a storage device. The sensor array includes but is not limited to OCT imaging sensors, visual evoked potential electrode sensors, robot end force / position sensors, eye movement tracking sensors, optical positioning tracking systems, etc.; the processor can be a central processing unit, a graphics processing unit, a multi-core processor, or an artificial intelligence chip, etc.; the storage device can be a distributed storage device or a centralized storage, which is not limited here.

[0024] In combination with the above explanations of the terms and the implementation environment, the application scenarios of the embodiments of the present application are described. The image-based ophthalmic robot control method provided in the embodiments of the present application can be applied in scenarios including but not limited to the following scenarios: In the field of minimally invasive surgery in ophthalmology, such as minimally invasive surgery for glaucoma, this technical solution also has significant application value. Although minimally invasive surgery has small trauma, the operation space is limited, and the operation precision is extremely high. During the operation, the multi-modal data fusion analysis result obtained by using this technical solution can help the doctor to more accurately locate the target region of the operation, such as the structure of the trabecular meshwork, and to monitor the interaction between the surgical instrument and the surrounding tissue in real time. Through dynamically generating robot operation correction control instructions, the robot can automatically adjust the motion trajectory and operation parameters of the surgical instrument according to the predicted visual function damage risk and postoperative visual recovery, ensuring the precision and safety of the operation. This not only helps to improve the success rate of the operation, but also reduces the occurrence of postoperative complications and promotes the rapid recovery of the patient's visual function after the operation.

[0025] Illustratively, the coal mine personnel safety situation dynamic perception method provided by the embodiments of the present application can also be applied to other application scenarios, and here is only used for illustration, and the specific application scenarios are not limited.

[0026] In an exemplary embodiment, as shown in Figure 1 An image-based ophthalmic robot control method is provided. The embodiment takes the method applied to the terminal in the foregoing implementation environment as an example. It can be understood that the method can also be applied to a server, and can also be applied to a system including a terminal and a server, and is realized through the interaction of the terminal and the server. The method includes the following steps 101 to 104: Step 101, acquiring an intraoperative three-dimensional OCT image sequence and a visual evoked potential signal, and segmenting a optic nerve region to generate an optic nerve spatial model.

[0027] Exemplarily, the intraoperative three-dimensional OCT image sequence is collected in real time by a high-resolution optical coherence tomography (OCT) device, which can provide high-precision anatomical structure information of the intraocular tissue, including detailed images of key structures such as the retina and optic nerve; at the same time, visual evoked potential signals are synchronously collected by professional electrophysiological monitoring equipment, which reflect the real-time functional state of the optic nerve during the operation and are an important indicator for evaluating whether the optic nerve is damaged. Specifically, the optic nerve region can be identified by using an improved U-Net segmentation network on the OCT image sequence: after inputting the original OCT image, multi-scale features are extracted by the encoder, spatial attention mechanism is embedded in the skip connection to enhance the optic disc edge feature response, and the decoder outputs the pixel-level segmentation mask. Further, the mask sequence is converted into a three-dimensional grid model of the optic nerve with topological connection relationship by using the Poisson surface reconstruction algorithm, which contains the node coordinates and normal vector information of the optic nerve fiber layer. For example, the segmentation network uses a multi-center data set to enhance generalization in the training stage, and adds a motion artifact adversarial training module to eliminate the interference of intraoperative shaking, and the generated spatial model of the optic nerve provides a spatial reference coordinate system for dynamic registration.

[0028] Step 102, according to the eye movement compensation algorithm, the optic nerve spatial model and the visual evoked potential signal are dynamically registered to generate a fusion data body.

[0029] Specifically, since the eyeball may move slightly during the operation, which may cause a spatio-temporal mismatch between the OCT image and the VEP signal. In order to eliminate this mismatch, the eye movement compensation algorithm can monitor the movement state of the eyeball in real time and make corresponding adjustments to the optic nerve spatial model. Exemplarily, this algorithm can be based on eye tracking technology, and by monitoring the movement trajectory of the eyeball, the optic nerve spatial model is dynamically corrected. Further, the compensated optic nerve spatial model is fused with the visual evoked potential signal to generate a fusion data body. The fusion data body integrates the anatomical structure and functional state information of the optic nerve, providing more comprehensive data support for the subsequent prediction model. For example, multi-modal data fusion techniques such as feature fusion or decision fusion based methods can be used to effectively integrate the optic nerve spatial model and the visual evoked potential signal. Through this dynamic registration and fusion process, the method can ensure the accuracy and consistency of the data, providing a reliable basis for subsequent prediction analysis.

[0030] Step 103, based on the fusion data body and the real-time operation parameters of the robot end instrument, the pre-trained neural network prediction model is used to calculate the visual function damage risk coefficient and the postoperative visual acuity recovery prediction value.

[0031] Specifically, the real-time operation parameters of the robotic end-effector instrument include the position, angle, force, and other information of the instrument, which directly affect the visual function status during the operation. By inputting the fused data volume and these real-time operation parameters into the pre-trained neural network prediction model, the model can calculate the visual function damage risk coefficient and postoperative visual acuity recovery prediction value based on a large amount of historical data and complex nonlinear relationships. For example, the neural network prediction model can use a deep learning architecture such as long short-term memory (LSTM) or convolutional neural network (CNN) to process time series data and spatial data. Further, the training process of the model can utilize a large amount of clinical data, including preoperative, intraoperative, and postoperative visual function evaluation results, to improve the accuracy and generalization ability of the model. For example, during the training process, data augmentation techniques such as random noise addition or data cropping can be used to improve the robustness of the model. Through this deep learning-based prediction method, the method can real-time assess the visual function risk during the operation and predict the postoperative visual acuity recovery, providing a scientific basis for the operation.

[0032] Step 104, dynamically generating robot operation correction control instructions according to the damage risk coefficient and the postoperative visual acuity recovery prediction value.

[0033] Specifically, according to the preset risk threshold and prediction target, the risk level and expected effect of the current operation can be real-time evaluated. If the damage risk coefficient exceeds the set threshold or the postoperative visual acuity recovery prediction value is lower than the expected target, the method can automatically adjust the operation parameters of the robotic end-effector instrument and generate correction control instructions. For example, the correction control instructions can include adjusting the motion trajectory, force, or operation speed of the instrument to reduce the risk of visual function damage and improve the possibility of postoperative visual acuity recovery. Further, these correction control instructions will be fed back to the robot control system in real time to ensure that the operation can be dynamically adjusted according to the real-time evaluation results. For example, if the prediction model finds that the current operation may cause excessive compression of the optic nerve, the force or position of the instrument can be automatically adjusted to avoid potential damage. Through this dynamic adjustment mechanism, the method can real-time optimize the operation and improve the safety and effectiveness of the operation.

[0034] The image-based ophthalmic robot control method, system, device and medium solve the problem of intraoperative optic nerve function dynamic monitoring, realize the quantitative correlation prediction of surgical operation parameters and postoperative visual recovery effect by establishing a real-time fusion analysis mechanism of anatomical structure and bioelectric signal, form a prognosis-oriented intraoperative closed-loop control, thereby improving the individualized precision of surgical operation and the protection effect of visual function.

[0035] In one embodiment, the optic nerve spatial model and the visual evoked potential signal are dynamically registered according to the eye movement compensation algorithm to generate a fused data volume, including: Performing optical flow field analysis on the three-dimensional OCT image sequence of adjacent frames to generate a tissue displacement vector field.

[0036] Illustratively, the optical flow field analysis is a computer vision technique used to estimate the direction and speed of pixel motion in image sequences. In this embodiment, by analyzing the three-dimensional OCT image sequence of adjacent frames, the displacement of intraocular tissues can be accurately calculated, thereby generating a tissue displacement vector field to capture the micro-movement of the eyeball during surgery and provide a basis for subsequent dynamic correction. Illustratively, classic optical flow algorithms such as Lucas-Kanade algorithm or Farneback algorithm can be used, which can effectively process three-dimensional image sequences and generate accurate displacement vector fields. Accurately quantify the real-time deformation of tissues caused by physiological tremor or instrument operation.

[0037] Constructing a dynamic correction matrix according to the tissue displacement vector field.

[0038] Specifically, the role of the dynamic correction matrix is to adjust the acquisition point coordinates of the visual evoked potential signal to keep consistent with the coordinate system of the optic nerve spatial model. Specifically, by matrix operation, the displacement information in the tissue displacement vector field is converted into coordinate adjustment parameters, thereby constructing the dynamic correction matrix. This process ensures the accurate alignment of the visual evoked potential signal and the optic nerve spatial model in space, providing a basis for subsequent data fusion. For example, the dynamic correction matrix can be constructed by linear transformation or nonlinear transformation, depending on the complexity and accuracy requirements of tissue displacement.

[0039] The collection point coordinates of the visual evoked potential signals are mapped to the corresponding nodes of the optic nerve spatial model through the dynamic correction matrix.

[0040] Specifically, after the adjustment of the dynamic correction matrix, the collection point coordinates of the visual evoked potential signals can be accurately mapped to the corresponding positions in the optic nerve spatial model, so as to spatially bind the functional signals with the anatomical structure, so that the visual evoked potential signals of each point can correspond to the specific position in the optic nerve spatial model. For example, the collection point coordinates can be accurately mapped to the nodes of the optic nerve spatial model through an interpolation algorithm or a nearest neighbor algorithm, to ensure the accuracy and reliability of the mapping.

[0041] The amplitude parameter and the latency offset parameter of the visual evoked potential signals are bound to the mapped nodes to generate a fusion data body.

[0042] Specifically, the amplitude parameter and the latency offset parameter of the visual evoked potential signals are bound to the corresponding nodes of the optic nerve spatial model to form a fusion data body containing anatomical structure and functional state. The amplitude parameter reflects the intensity of the optic nerve signal, and the latency offset parameter reflects the delay of the signal, both of which are important indicators for evaluating the functional state of the optic nerve. By binding these parameters to the nodes of the optic nerve spatial model, the generated fusion data body can provide comprehensive information of the intraocular tissue, providing more abundant data support for subsequent analysis and prediction. For example, the amplitude parameter can be determined by the peak value of the signal, and the latency offset parameter can be calculated by the difference between the starting time of the signal and the standard time. Through this binding method, the fusion data body not only contains the anatomical structure of the optic nerve, but also contains the real-time information of its functional state, providing more comprehensive guidance for surgical operation. The above embodiments can realize accurate dynamic registration and fusion of intraoperative optic nerve bioelectric signals and multi-modal images, improve the safety and effectiveness of the operation, and also provide a data basis for real-time evaluation of the functional state of the optic nerve, enhancing the postoperative visual function protection effect.

[0043] In one of the embodiments, based on the fusion data body and the real-time operation parameters of the robot end instrument, the visual function damage risk coefficient and the postoperative visual acuity recovery prediction value are calculated through a pre-trained neural network prediction model, including: Based on the amplitude parameter, the latency offset parameter, and the real-time pressure change rate of the robot end instrument, a dynamic optic nerve function index is calculated.

[0044] Specifically, the core bioelectric parameters in the fusion data body are continuously acquired during the operation, including the amplitude parameter of the visual evoked potential signal (reflecting the synchronous discharge intensity of the neural cell group) and the latency shift parameter (indicating the change in neural conduction velocity), while the pressure rate of change of the robot end instrument in the XYZ axial direction (characterizing the instantaneous mechanical stimulation intensity of the instrument on the tissue) is extracted from the six-dimensional force sensor. Based on the above three groups of dynamic parameters, a dynamic optic nerve function index is constructed, including: normalizing the amplitude parameter to eliminate individual potential amplitude base differences; calculating the first derivative of the latency shift to capture the deterioration trend of neural conduction block; and nonlinearly weighting the pressure rate of change and the aforementioned bioelectric parameters to form a scalar index that comprehensively reflects the real-time functional status of the optic nerve. This index can sensitively identify the decline in neural conduction efficiency caused by optic disc compression in glaucoma surgery, such as the abnormal increase in latency derivative when the pressure rate of change increases sharply. Further, the dynamic optic nerve function index can be calculated by weighted summation, in which the amplitude parameter, the latency shift parameter, and the real-time pressure rate of change are respectively assigned different weights to reflect their importance in the evaluation of optic nerve function. For example, the weight of the amplitude parameter can be set to 0.4, the weight of the latency shift parameter can be set to 0.3, and the weight of the real-time pressure rate of change can be set to 0.3, thereby obtaining a comprehensive dynamic optic nerve function index for subsequent prediction analysis.

[0045] The dynamic optic nerve function index and the real-time pressure are input into a pre-trained convolutional recurrent neural network, and the output is a visual function damage risk coefficient.

[0046] Specifically, the convolutional recurrent neural network (CRNN) model pre-trained in this embodiment undertakes the core prediction task. The convolutional recurrent neural network (Convolutional Recurrent Neural Network, CRNN) is a deep learning architecture that combines the characteristics of convolutional neural network (CNN) and recurrent neural network (RNN). It combines the advantages of CNN in processing spatial data and the advantages of RNN in processing time series data, and can effectively process complex data containing both spatial information and time information. In this application, CRNN is used to process the spatial features in the fusion data body and the time series features of the real-time operation parameters of the robot end instrument, to realize accurate prediction of the risk of visual function damage. Exemplarily, in this embodiment, the input layer of the network is designed as a dual-channel architecture, in which: the first channel receives the time series flow data of the DNFI, extracts the feature decay pattern through two layers of long short-term memory (LSTM) units, and pays special attention to the gradient mutation features within a 15ms time window; the second channel inputs the spatial distribution map of the instrument operation pressure, and adopts a three-dimensional convolution kernel to scan the pressure conduction hot spot on the optic nerve spatial model. After the dual-channel output is fused through the attention gate mechanism, the visual function damage risk coefficient (0-1 continuous value) is output through the regression layer. During model training, a multi-center surgery data set is used to enhance robustness, and a gradient penalty mechanism is introduced to prevent overfitting, ensuring that in complex scenarios such as diabetic retinopathy, high-risk operation modes (such as when the ultrasonic emulsification probe approaches the optic nerve, the DNFI decays exponentially) can still be identified stably.

[0047] Based on the functional relationship between the visual function damage risk coefficient and the preoperative baseline visual acuity, the postoperative visual acuity recovery prediction value is calculated.

[0048] Specifically, the postoperative visual recovery prediction is achieved based on the dynamic correlation of the damage risk and the preoperative baseline. Exemplarily, by reading the key parameters in the preoperative visual function profile of the patient, including the mean defect (MD) value of static visual field and the baseline amplitude of pattern visual evoked potential P100 wave, the real-time damage risk coefficient during the operation is input into the prognosis function model: a bivariate interaction equation is established to consider the buffering effect of preoperative optic nerve compensation on intraoperative damage; a time integral operation is introduced to accumulate the exposure time of high-risk operation in the entire surgical stage; and a postoperative visual recovery prediction value (Snellen visual acuity percentage) is output. This mechanism can predict the prognosis difference of different dissection paths in macular surgery, for example, when the cumulative risk value exceeds the compensation threshold, it automatically prompts to choose a temporal approach to avoid central foveal function damage. This embodiment constructs a high-sensitivity functional index by dynamically weighting and fusing multiple source parameters, analyzes the complex mapping relationship between mechanical stimulation and neural response by using a double-channel spatiotemporal joint modeling, and realizes individualized prognosis prediction based on preoperative and intraoperative parameter coupling algorithm. In clinical application, the visual function evaluation time point can be advanced from postoperative to intraoperative decision-making link, so that the doctor can adjust the surgical procedure according to the prediction value during the key operation stage such as retinal peeling, thereby reducing the risk of irreversible optic nerve damage.

[0049] In one of the embodiments, a robot operation correction control instruction is dynamically generated according to the damage risk coefficient and the postoperative visual recovery prediction value, including: Based on the real-time position coordinates of the robot end instrument in the three-dimensional OCT image sequence, the current surgical stage identifier is generated by a preset surgical stage division rule.

[0050] Specifically, based on the real-time spatial coordinates of the robot end instrument in the three-dimensional OCT image sequence (obtained by registering the instrument retroreflective marker point and the OCT voxel coordinates), the stage intelligent division is performed in combination with the surgical procedure knowledge base. For example, the preset surgical stage division rule can include multi-dimensional condition judgment: when the instrument tip enters the space range of 0.5 mm from the optic disc, it is marked as "optic nerve adjacent stage"; when the change rate of vitreous proliferation membrane traction force detected exceeds 0.3 N / s, it is marked as "proliferation membrane peeling stage". These spatial and mechanical parameters jointly constitute the stage division basis to generate a surgical stage identifier with clinical semantics.

[0051] It is judged whether the current surgical stage identifier belongs to the preset high-risk stage set.

[0052] Exemplarily, when judging the risk attribute of the stage, a preset high-risk stage set can be called for matching comparison, which is dynamically configured according to the type of surgery before surgery: for example, in the surgery of diabetic retinopathy, "epipapillary membrane peeling" and "macular pre-membrane hooking" are included in the high-risk set. The matching mechanism adopts a double-verification strategy, for example, first screening through the positional relationship between the instrument coordinates and the anatomical partition of the optic nerve, and then combining the current pressure fluctuation spectrum characteristics (such as the energy increase of the 0.5-2Hz frequency band) for secondary confirmation. When the spatial positioning condition and the mechanical characteristic condition are met at the same time, it is determined that there is a risk of mechanical damage to the optic nerve in the current stage.

[0053] If the current surgery stage identifier belongs to the preset high-risk stage set, the maximum allowable movement speed and the pressure safety threshold are calculated according to the visual function damage risk coefficient.

[0054] Specifically, based on the real-time visual function damage risk coefficient, the mechanical operation limit parameters are calculated by a non-linear mapping function, wherein the calculation formula of the maximum allowable movement speed is: ; Wherein, is the maximum allowable movement speed, is the current surgery stage reference speed, which is obtained from the preset initial speed of safe operation of the instrument determined by the preset surgical knowledge base, is the instrument attenuation coefficient, which represents the adjustment factor of the sensitivity of different instruments to risk, is the visual function damage risk coefficient; The calculation formula of the pressure safety threshold is: ; Wherein, is the pressure safety threshold, is the critical pressure stress of the optic nerve (for example, glaucoma patients = 25mN), is the cumulative damage factor (positively correlated with the preoperative cup-to-disc ratio), is the duration of the current stage.

[0055] If the postoperative visual acuity recovery prediction value is less than the preset recovery threshold, a motion pause instruction is generated and an alarm signal is triggered.

[0056] Specifically, the postoperative visual acuity recovery prediction value is calculated based on the visual function impairment risk coefficient and the preoperative baseline visual acuity, which reflects the possibility of postoperative visual acuity recovery of the patient. The preset recovery threshold is preset according to clinical experience and surgical goals, and is used to judge whether the postoperative visual acuity recovery reaches the expected goal. If the postoperative visual acuity recovery prediction value is lower than the preset recovery threshold, it means that the current surgical operation may cause poor postoperative visual acuity recovery. In this case, a motion pause instruction is generated to immediately stop the motion of the robot end instrument to prevent further damage; at the same time, an alarm signal is triggered to remind the surgeon to pay attention to the risk of the current operation, so as to take timely measures for adjustment. Through this mechanism, the method can intervene in time when the postoperative visual acuity recovery is poor, and establishes an active protection paradigm for ophthalmic robot surgery, ensuring the safety and effectiveness of the surgery.

[0057] In one of the embodiments, the dynamic correction matrix is constructed according to the tissue displacement vector field, further comprising: The affine transformation of the tissue displacement vector field is modeled to generate a spatial compensation parameter matrix, and the spatial compensation parameter matrix includes translation, rotation and scaling parameters; The original coordinates of the visual evoked potential signal acquisition points are subjected to spatial transformation operation by using the spatial compensation parameter matrix to generate corrected spatial coordinates; According to the Euclidean distance between the corrected spatial coordinates and the nodes in the optic nerve spatial model, an effective mapping node set is screened; Based on the effective mapping node set, the spatial coordinate distribution of the fusion data body is updated.

[0058] Exemplarily, the embodiment constructs a dynamic correction matrix according to the in-situ collected tissue displacement vector field, and realizes dynamic registration of the visual evoked potential signal and the anatomical structure. The method firstly models the affine transformation of the tissue displacement vector field, which is derived from the optical flow field analysis result of the three-dimensional OCT image sequence and reflects the real-time spatial deformation of the living tissue. Specifically, the singular value decomposition algorithm can be used to fit the optimal geometric transformation relationship, and a spatial compensation parameter matrix containing the translation vector, rotation matrix and scaling factor is generated. The matrix quantifies the spatial drift of the optic nerve caused by the compression of the surgical instrument or the breathing fluctuation, for example, the rigid displacement component caused by the retina traction can be accurately compensated in the macular surgery. Further, the spatial compensation parameter matrix is used to perform spatial transformation operation on the original coordinates of the visual evoked potential signal acquisition point. The original coordinates are provided by the optical positioning system of the scalp electrode, and the acquisition point is mapped to the OCT image coordinate system through homogeneous coordinate conversion, affine transformation calculation is performed to generate corrected spatial coordinates, and double-precision floating-point operation is used to ensure spatial accuracy to eliminate the electrode position deviation caused by in-situ medium disturbance, for example, to eliminate the 0.2 millimeter level coordinate offset caused by the ultrasonic emulsification probe vibration, and to improve the spatial positioning accuracy of the bioelectric signal. The method selects an effective mapping node set according to the Euclidean distance between the corrected spatial coordinates and the nodes in the optic nerve spatial model. The processing process introduces a weighted distance measurement mechanism to narrow the distance threshold for anatomical regions with significant curvature characteristics. The nearest neighbor search is accelerated by establishing a spatial topological index tree, and a node cluster with a distance error less than a preset threshold is selected. The operation realizes high-precision spatial association between the signal and the anatomical structure, for example, effectively eliminates the non-physiological mapping points in the optic disc edge area, and reduces the mismatch rate to one third of the traditional method. The spatial coordinate distribution of the fusion data body is updated based on the effective mapping node set, the node coordinate attribute of the reconstructed optic nerve spatial model is written into the heterogeneous data storage structure, and the spatial confidence evaluation index is established, and a dynamic weight coefficient is added to the displacement compensated node. This process improves the spatial consistency of multi-source data, for example, maintains the real-time matching between the retinal nerve fiber layer and the evoked potential during the retinal peeling stage, and ensures that the anatomical basis of functional damage assessment is always accurate and reliable.

[0059] In one embodiment, the training method of the convolutional recurrent neural network comprises: constructing a training data set, the training data set containing preoperative three-dimensional OCT image features, intraoperative instrument operation parameter sequences and postoperative clinical visual function evaluation results; generating instrument operation parameter and visual function damage correlation simulation data through the adversarial generative network, the correlation simulation data being used to expand the training data set; using a long short-term memory module to extract the time sequence dependent features of the instrument operation parameters, and using an attention mechanism to weight and fuse the time sequence dependent features and the image features; The network weight parameters are optimized by a back propagation algorithm, so that the KL divergence between the risk coefficient of visual function impairment output by the convolutional recurrent neural network and the postoperative clinical evaluation result converges to a set range.

[0060] Specifically, preoperative three-dimensional OCT image features are obtained by high-resolution optical coherence tomography (OCT) equipment, which can provide detailed anatomical structure information of intraocular tissues; intraoperative instrument operation parameter sequence records real-time parameters such as position, angle, and force of the instrument during the operation, which directly affect the operation effect. Postoperative clinical visual function evaluation results are obtained through clinical examination, reflecting the actual situation of the patient's postoperative visual recovery. By integrating these multi-dimensional data, the training data set can provide comprehensive information support for network training. Further, the correlation simulation data of instrument operation parameters and visual function damage are generated by the generative adversarial network (GAN), which is a generative adversarial model that can learn the distribution of data and generate new data samples. In this embodiment, GAN is used to generate simulation data with similar distribution to real data, which can supplement the training data set and improve the generalization ability of the network to different situations. Specifically, the generator network in GAN generates simulation data according to the input noise, while the discriminator network tries to distinguish real data and simulation data. Through the adversarial process between the generator and the discriminator, the generator can generate more and more realistic simulation data, which can be used to expand the training data set and increase the diversity and quantity of data. Exemplarily, a long short-term memory module (LSTM) is used to extract the time-dependent features of the instrument operation parameters, and the time-dependent features and image features are weighted and fused through an attention mechanism. The long short-term memory module is a recurrent neural network structure that can effectively handle long-term dependencies in time series data. In this embodiment, LSTM is used to extract the time-dependent features of the intraoperative instrument operation parameter sequence, which reflect the variation law of the instrument operation parameters over time. At the same time, through the attention mechanism, the network can automatically learn the weight relationship between the image features and the time-dependent features, and weightedly fuse the two features, so as to more comprehensively capture the key information in the operation process. For example, the attention mechanism can dynamically adjust the weight of the features according to the importance of different time steps and the importance of different image regions, so that the network can more effectively utilize these features for prediction. Further, the network weight parameters are optimized by the back propagation algorithm to make the KL divergence of the visual function damage risk coefficient output by the convolutional recurrent neural network and the postoperative clinical evaluation results converge within a set range. The back propagation algorithm is a commonly used neural network training algorithm that calculates the gradient of the loss function with respect to the network weights to update the network weights to minimize the loss function. In this embodiment, the loss function uses KL divergence, i.e., Kullback-Leibler divergence, which measures the difference between two probability distributions. By optimizing the network weights, the KL divergence between the probability distribution of the visual function damage risk coefficient output by the network and the probability distribution of the postoperative clinical evaluation results converges within a set range, thereby improving the accuracy and reliability of the network prediction.For example, the threshold of KL divergence can be set as 0.1, when the KL divergence is less than the threshold, it is considered that the network training reaches convergence, and the network can accurately predict the visual function damage risk coefficient.

[0061] In one of the embodiments, the image-based ophthalmic robot control method further comprises: In the postoperative verification stage, the visual function damage risk coefficient sequence recorded in the historical operation process and the postoperative visual recovery prediction value of the key operation point are obtained; The visual function damage risk coefficient sequence and the postoperative visual recovery prediction value are subjected to Gaussian mixture model clustering analysis to generate a high-risk operation mode feature set; Based on the high-risk operation mode feature set and the anatomical structure features of the preoperative three-dimensional OCT image, a prognosis optimization score report is generated through a random forest algorithm, and the prognosis optimization score report includes operation path planning suggestions and risk avoidance strategies.

[0062] Exemplarily, the embodiment establishes a dynamic correlation between high-risk operation modes and anatomical features through multi-dimensional data analysis in the postoperative verification stage, and generates a quantitative evaluation report that can guide improvement. The method first obtains a sequence of visual function damage risk coefficients recorded during the historical surgery process and postoperative visual acuity recovery prediction values at key operation points. Specifically, the sequence of risk coefficients is derived from the convolutional recurrent neural network prediction results every 50 milliseconds during the surgery, covering instrument movement trajectory, pressure gradient, and biological electrical signal mutation events. The postoperative visual acuity recovery prediction values are generated by a coupling model of preoperative baseline visual acuity and real-time risk values during the surgery, for example, the correlation parameters in macular surgery include foveal thickness change value and visual acuity loss. Gaussian mixture model clustering analysis is performed on the above data. Exemplarily, the expectation maximization algorithm can be used to iteratively optimize the model parameters: first, determine the optimal number of clusters according to the Bayesian information criterion, then identify the spatiotemporal distribution characteristics of each cluster through covariance matrix decomposition, and also introduce an information entropy weight adjustment mechanism to give higher clustering weight to operation intervals that exceed the risk threshold for more than 500 milliseconds. The generated high-risk operation mode feature set includes three core dimensions: risk coefficient fluctuation spectrum characteristics, instrument path deviation mode, and predicted value decay curve morphology, for example, identifying the typical high-risk mode of "ultrasonic emulsification stage pressure oscillation accompanied by optic nerve conduction delay". Random forest is an ensemble learning algorithm that can handle a large number of features and provide accurate classification or regression results. In this embodiment, the random forest algorithm combines the high-risk operation mode feature set and the anatomical structure features of the preoperative three-dimensional OCT image to evaluate each surgical case. The prognosis optimization score report includes surgery path planning suggestions and risk avoidance strategies, which aim to optimize the surgery path and reduce the occurrence of high-risk operation modes, thereby improving the safety of the surgery and the possibility of postoperative visual acuity recovery. For example, the report may suggest using more cautious operation parameters near specific anatomical structures, or adjusting the surgery path to avoid high-risk areas.

[0063] In summary, the image-based ophthalmic robot control method provided in the present application constructs a dynamic optic nerve spatial model by fusing three-dimensional OCT images and visual evoked potential signals in real time during the surgery, and realizes millisecond-level precise registration of biological electrical signals and anatomical structures based on eye movement compensation algorithm; on the basis of this fusion data, the pre-trained neural network model is used to calculate the visual function damage risk coefficient and the postoperative visual acuity recovery prediction value in real time combined with the robot operation parameters, and the instrument movement constraint instructions and risk avoidance strategies are dynamically generated according to the prediction results. This technical solution innovatively establishes a real-time prediction correlation between intraoperative operation parameters and postoperative visual acuity recovery effect, actively optimizes the instrument path and operation force during the surgery, not only improves the precision of individualized surgical plan, but also realizes the paradigm shift from traditional anatomical positioning to optic nerve function protection, and fundamentally solves the core technical bottleneck of being unable to dynamically evaluate and protect the nerve function in ophthalmic robot surgery.

[0064] It should be understood that although the steps in the flowcharts involved in the embodiments described above are shown in sequence according to the arrows, these steps are not necessarily executed in the order indicated by the arrows. Unless otherwise specified herein, the execution of these steps is not strictly limited in sequence, and these steps can be executed in other orders. Moreover, at least some of the steps in the flowcharts involved in the embodiments described above can include multiple steps or multiple stages, which are not necessarily executed at the same time but can be executed at different times, and the execution of these steps or stages is not necessarily sequential but can be executed alternately or alternately with at least some of the other steps or steps or stages in other steps.

[0065] Based on the same inventive concept, the embodiments of the present application also provide an image-based ophthalmic robot control system 10 for implementing the above-mentioned image-based ophthalmic robot control method. The implementation scheme for solving the problem provided by the system is similar to the implementation scheme described in the above method, so the specific limitations in one or more image-based ophthalmic robot control system 10 embodiments provided below can refer to the limitations of the image-based ophthalmic robot control method described above, which will not be repeated here.

[0066] In one exemplary embodiment, as shown in Figure 2 An image-based ophthalmic robot control system 10 is provided, comprising: An image segmentation module 11 is configured to acquire an intraoperative three-dimensional OCT image sequence and a visual evoked potential signal, segment a optic nerve region to generate an optic nerve spatial model; A data fusion module 12 is configured to perform dynamic registration on the optic nerve spatial model and the visual evoked potential signal according to an eye movement compensation algorithm to generate a fusion data body; A visual function prediction analysis module 13 is configured to calculate a visual function damage risk coefficient and a postoperative visual acuity recovery prediction value based on the fusion data body and real-time operation parameters of a robot end instrument through a pre-trained neural network prediction model; A control instruction generation module 14 is configured to dynamically generate robot operation correction control instructions according to the damage risk coefficient and the postoperative visual acuity recovery prediction value.

[0067] In one embodiment, the data fusion module 12 comprises: An optical flow analysis unit is configured to perform optical flow field analysis on the three-dimensional OCT image sequence of adjacent frames to generate a tissue displacement vector field; An instruction execution unit is configured to construct a dynamic correction matrix according to the tissue displacement vector field; A coordinate mapping unit is configured to map coordinates of the collected points of the visual evoked potential signals to corresponding nodes of the optic nerve spatial model through a dynamic correction matrix. A data binding unit is configured to bind the amplitude parameter and the latency offset parameter of the visual evoked potential signals to the mapped nodes to generate a fusion data body.

[0068] In one of the embodiments, the visual function prediction analysis module 13 comprises: A function index calculation unit is configured to calculate a dynamic optic nerve function index based on the amplitude parameter, the latency offset parameter, and a real-time pressure change rate of the robot end instrument. A risk prediction unit is configured to input the dynamic optic nerve function index and the real-time pressure into a pre-trained convolutional recurrent neural network to output a visual function damage risk coefficient. A visual acuity prediction unit is configured to calculate a postoperative visual acuity recovery prediction value based on a functional relationship between the visual function damage risk coefficient and the preoperative baseline visual acuity.

[0069] In one of the embodiments, the control instruction generation module 14 comprises: A stage identification unit is configured to generate a current surgical stage identifier based on real-time position coordinates of the robot end instrument in the three-dimensional OCT image sequence through a preset surgical stage division rule. A risk judgment unit is configured to determine whether the current surgical stage identifier belongs to a preset high-risk stage set. A threshold calculation unit is configured to calculate a maximum allowable movement speed and a pressure safety threshold value according to the visual function damage risk coefficient if the current surgical stage identifier belongs to the preset high-risk stage set. An instruction execution unit is configured to generate a motion pause instruction and trigger an alarm signal if the postoperative visual acuity recovery prediction value is less than a preset recovery threshold value.

[0070] In one of the embodiments, the instruction execution unit can be further configured to perform the following steps: An affine transformation modeling is performed on the tissue displacement vector field to generate a spatial compensation parameter matrix, which includes translation, rotation, and scaling parameters. A spatial transformation operation is performed on the original coordinates of the visual evoked potential signal collection points using the spatial compensation parameter matrix to generate corrected spatial coordinates. According to the Euclidean distance between the corrected spatial coordinates and the nodes in the optic nerve spatial model, an effective mapping node set is screened. The spatial coordinate distribution of the fusion data body is updated based on the effective mapping node set.

[0071] In one of the embodiments, the training method of the convolutional recurrent neural network in the risk prediction unit comprises: constructing a training data set, the training data set containing preoperative three-dimensional OCT image features, intraoperative instrument operation parameter sequences, and postoperative clinical visual function evaluation results; generating instrument operation parameter and visual function damage correlation simulation data by the generative adversarial network, the correlation simulation data being used to expand the training data set; extracting the time sequence dependent features of the instrument operation parameters by using the long short-term memory module, and weighting and fusing the time sequence dependent features and the image features by using the attention mechanism; optimizing the network weight parameters by using the back propagation algorithm, so that the KL divergence between the visual function damage risk coefficients output by the convolutional recurrent neural network and the postoperative clinical evaluation results converges to a set range.

[0072] In one of the embodiments, the image-based ophthalmic robot control system 10 further comprises a postoperative optimization unit for performing the following steps: In the postoperative verification stage, the visual function damage risk coefficient sequence recorded in the historical surgery process and the postoperative visual recovery prediction value of the key operation point are obtained; performing Gaussian mixture model clustering analysis on the visual function damage risk coefficient sequence and the postoperative visual recovery prediction value to generate a high-risk operation mode feature set; based on the high-risk operation mode feature set and the anatomical structure features of the preoperative three-dimensional OCT image, generating a prognosis optimization score report by using the random forest algorithm, the prognosis optimization score report including surgery path planning suggestions and risk avoidance strategies.

[0073] In one embodiment, a computer device is provided, comprising a memory and a processor, the memory storing a computer program, and the processor implementing the steps of the image-based ophthalmic robot control method as described above when executing the computer program.

[0074] In one embodiment, a computer readable storage medium is provided, which stores a computer program, and the computer program is executed by a processor to implement the steps of the above method embodiments.

[0075] For the device embodiment, since it basically corresponds to the method embodiment, the relevant parts are described in the part of the method embodiment. The device embodiments described above are only illustrative, and the components described as separate components can be or can not be physically separated, and the components displayed as units can be or can not be physical units, that is, they can be located in one place, or can be distributed on multiple network units. According to actual needs, part or all of the modules can be selected to achieve the purpose of the present disclosure. Those skilled in the art can understand and implement without creative labor.

[0076] The above-described embodiments only express several implementation manners of the application, the description is more specific and detailed, but it cannot be understood as the limitation of the patent scope of the application. It should be pointed out that for ordinary skilled in the art, without departing from the concept of the application, several modifications and improvements can be made, which are within the protection scope of the application.

Claims

1. An image-based ophthalmic robot control method, characterized in that: The method comprises: Acquire intraoperative 3D OCT image sequences and visual evoked potential signals, segment the optic nerve area and generate an optic nerve spatial model; Dynamically registering the optic nerve space model with the visual evoked potential signal according to an eye movement compensation algorithm to generate a fused data volume; Based on the fused data volume and the real-time operating parameters of the robot end instrument, a pre-trained neural network prediction model is used to calculate the risk coefficient of visual function damage and the predicted value of postoperative vision recovery; Dynamically generate robot operation correction control instructions based on the injury risk coefficient and the postoperative vision recovery prediction value.

2. The method according to claim 1, characterized in that The dynamically registering the optic nerve space model with the visual evoked potential signal according to the eye movement compensation algorithm to generate a fused data volume includes: Optical flow analysis is performed on adjacent frames of the 3D OCT image sequence to generate a tissue displacement vector field; constructing a dynamic correction matrix according to the tissue displacement vector field; Mapping the acquisition point coordinates of the visual evoked potential signal to corresponding nodes of the optic nerve space model through the dynamic correction matrix; The amplitude parameter and the latency offset parameter of the visual evoked potential signal are bound to the mapped nodes to generate the fused data volume.

3. The method according to claim 2, characterized in that The method of calculating the risk coefficient of visual function damage and the predicted value of postoperative vision recovery by a pre-trained neural network prediction model based on the fused data volume and the real-time operating parameters of the robot end instrument includes: Calculating a dynamic optic nerve function index based on the amplitude parameter, the latency offset parameter, and the real-time pressure change rate of the robot end instrument; Inputting the dynamic optic nerve function index and real-time pressure into a pre-trained convolutional recurrent neural network to output the visual function damage risk coefficient; Based on the functional relationship between the visual function impairment risk coefficient and the preoperative baseline visual acuity, the postoperative visual acuity recovery prediction value is calculated.

4. The method according to claim 1, wherein The dynamically generating robot operation correction control instructions according to the injury risk coefficient and the postoperative vision recovery prediction value includes: Based on the real-time position coordinates of the robot end instrument in the three-dimensional OCT image sequence, a current surgical stage identifier is generated according to a preset surgical stage division rule; Determining whether the current surgical stage identifier belongs to a preset high-risk stage set; If the current surgical stage identifier belongs to the preset high-risk stage set, calculating the maximum allowable movement speed and pressure safety threshold according to the visual function damage risk coefficient; If the predicted value of postoperative vision recovery is less than a preset recovery threshold, a movement pause instruction is generated and an alarm signal is triggered.

5. The method according to claim 2, characterized in that The constructing of a dynamic correction matrix according to the tissue displacement vector field further includes: Performing affine transformation modeling on the tissue displacement vector field to generate a spatial compensation parameter matrix, wherein the spatial compensation parameter matrix includes translation, rotation, and scaling parameters; Performing a spatial transformation operation on the original coordinates of the visual evoked potential signal acquisition points using the spatial compensation parameter matrix to generate corrected spatial coordinates; screening a valid mapping node set according to the Euclidean distance between the corrected spatial coordinates and the nodes in the optic nerve space model; The spatial coordinate distribution of the fused data volume is updated based on the valid mapping node set.

6. The method according to claim 3, characterized in that The training method of the convolutional recurrent neural network includes: Constructing a training data set, wherein the training data set includes preoperative three-dimensional OCT image features, intraoperative instrument operation parameter sequences, and postoperative clinical visual function evaluation results; generating simulation data on the correlation between instrument operating parameters and visual function impairment by using a generative adversarial network, wherein the simulation data on the correlation is used to expand the training data set; A long short-term memory module is used to extract the temporal dependency features of the instrument operation parameters, and the temporal dependency features are weightedly fused with the image features through an attention mechanism; The network weight parameters are optimized by the back propagation algorithm so that the KL divergence between the visual function impairment risk coefficient output by the convolutional recurrent neural network and the postoperative clinical evaluation results converges to a set range.

7. The method according to claim 6, characterized in that The method further comprises: During the postoperative verification phase, the risk coefficient sequence of visual function damage and the predicted value of postoperative visual recovery at key operation points recorded during the historical surgical process were obtained; Performing a Gaussian mixture model cluster analysis on the visual function impairment risk coefficient sequence and the postoperative vision recovery prediction value to generate a high-risk operation mode feature set; Based on the high-risk operation mode feature set and the anatomical structure features of the preoperative three-dimensional OCT image, a prognosis optimization score report is generated by a random forest algorithm. The prognosis optimization score report includes surgical path planning suggestions and risk avoidance strategies.

8. An image-based ophthalmic robot control system, characterized in that: The system comprises: Image segmentation module, used to obtain intraoperative 3D OCT image sequences and visual evoked potential signals, segment the optic nerve area and generate an optic nerve spatial model; a data fusion module, configured to dynamically register the optic nerve spatial model with the visual evoked potential signal according to an eye movement compensation algorithm to generate a fused data volume; A visual function prediction and analysis module, configured to calculate the visual function impairment risk coefficient and the postoperative vision recovery prediction value through a pre-trained neural network prediction model based on the fused data volume and the real-time operating parameters of the robot end instrument; A control instruction generation module is used to dynamically generate robot operation correction control instructions based on the injury risk coefficient and the postoperative vision recovery prediction value.

9. A computer device comprising a memory and a processor, wherein the memory stores a computer program, wherein: When the processor executes the computer program, the method according to any one of claims 1 to 7 is implemented.

10. A computer-readable storage medium having a computer program stored thereon, characterized in that: When the computer program is executed by a processor, the method according to any one of claims 1 to 7 is implemented.

Citation Information

Patent Citations

  • Ophthalmologic operation real-time navigation system and method based on dynamic visual field tracking

    CN119818291A

  • Intraoperative image-guided tools for ophthalmic surgery

    US20220346884A1

Cited By

  • Vision-electrophysiology multi-mode-based eye movement tracking zoom control system and zoom glasses

    CN122331147A