Systems and methods for determining the location of a patient's gross target volume.
By modeling the motion of target volume lesions during radiotherapy, and using two-dimensional image data and adaptive motion models to predict lesion locations, the problem of uncertain target volume locations is solved, achieving more precise energy delivery and more efficient treatment results.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- SIEMENS HEALTHINEERS AG
- Filing Date
- 2025-12-24
- Publication Date
- 2026-06-30
AI Technical Summary
In traditional radiotherapy, the patient's anatomical structure and tumor lesions make it difficult to accurately determine the target volume location due to respiration and visceral movement, resulting in inaccurate energy delivery and affecting the treatment effect.
By modeling the motion of lesions in a large target volume, multiple two-dimensional image data and an adaptive motion model are used to predict the future location of the lesion, generating control signals to adjust the operation of the radiotherapy equipment to accurately target the target volume.
It improves the accuracy of energy delivery during radiotherapy, reduces computational resource consumption, shortens equipment downtime, and improves treatment efficiency and effectiveness.
Smart Images

Figure CN122297931A_ABST
Abstract
Description
Technical Field
[0001] This application generally relates to systems and methods for predicting the location of a gross target volume within a patient during radiotherapy, and in some embodiments relates to systems and methods for predicting the location of a gross target volume within a patient during radiotherapy by modeling the movement of one or more lesions of the gross target volume over time. Background Technology
[0002] Radiation therapy (also known as "radiotherapy" or "RT") involves delivering radiation (energy) to a target within a patient's body during treatment. For example, a linear accelerator (LINAC) can be configured to move relative to the patient to multiple control points according to an RT treatment plan and deliver energy to target tissues (such as tumors, lesions, etc.) to treat cancer located within the patient's body. This allows energy to be delivered specifically to the target tissue with the aim of reducing or eliminating the target tissue without affecting surrounding tissues.
[0003] However, traditional methods for implementing treatment plans for patients are difficult to execute in a controlled manner. For example, a patient's anatomy, tumor lesions, etc., can move during respiration due to reflexive movement and / or due to visceral movement (the natural movement of organs within the body). This can lead to displacement of the gross target volume (GTV) being targeted during RT therapy and misalignment relative to LINAC. To reduce the likelihood of this, the patient being treated can be instructed to hold their breath while energy is delivered (often referred to as deep inspiratory breath-hold or "DIBH"). However, the success rate of this technique can vary drastically and cannot fully account for unconscious or visceral movements. Furthermore, other techniques involving monitoring patient movement based on external features (such as surrogate devices) may suffer from "drift," leading to inaccurate determination of GTV location. Therefore, the energy delivered to the patient's GTV may be suboptimal, as some of the energy originally intended to target the GTV may be misaligned and delivered to different parts of the GTV or one or more organs at risk (OARs). Summary of the Invention
[0004] For the reasons mentioned above, there is a need for a system and method for predicting the location of a gross target volume within a patient during radiotherapy by modeling the movement of one or more lesions of the gross target volume over time.
[0005] The methods and systems discussed in this paper aim to address the challenge of predicting lesion location in 3D space during radiotherapy, where the lesion location may change at least to some extent due to patient insensitivity, reflexes, or visceral movement. More specifically, the currently disclosed techniques aim to model the motion of the lesion's gastrointestinal tract (GTV) during radiotherapy, thereby minimizing the deviation between the lesion's predetermined location (e.g., posture) in 3D space and its actual location (as these locations may change).
[0006] In some embodiments, a system including one or more processors can be configured to acquire image data associated with multiple two-dimensional (2D) images generated by imaging devices positioned around a patient. These systems can determine the location of one or more lesions of the patient at a first time point based on the multiple 2D images. The system can then determine the future location of the one or more lesions at a later time point based on the location of the one or more lesions and trajectories representing the expected movement of the one or more lesions. In some examples, the system can then generate control signals to control the operation of one or more devices. For example, the system can generate control signals during RT treatment to control the operation of a LINAC, wherein such control signals adjust the position of one or more components of the LINAC. In other examples, the system can generate control signals to generate a beam direction view of the predicted location of the GTV over time.
[0007] Several technological advantages can be achieved by implementing some or all of the techniques described herein. First, the location of one or more lesions in 3D space can be determined more precisely compared to conventional methods. This, in turn, can improve treatment by delivering energy more accurately during RT therapy. Furthermore, systems implementing the techniques described herein can account for unpredictable uncertainties arising from involuntary patient movements, reflexive movements, or visceral movements. In addition to improved energy delivery, these advantages can reduce the consumption of computational resources. For example, systems including processors and memory can be completely preserved or retained, otherwise dedicated to determining the location of the GTV (e.g., based on relative motion of surrogates). Moreover, medical devices such as LINAC can be configured to operate with less downtime between instances of delivered energy, thus delivering energy to the GTV more quickly, which can improve the effectiveness of RT therapy (e.g., targeting and destroying lesion cells). These advantages, through radiotherapy and through improved energy delivery, collectively promote more effective and efficient cancer treatment.
[0008] In one embodiment, a system is disclosed for predicting the location of a gross target volume within a patient during radiotherapy by modeling the motion of one or more lesions of a gross target volume over time. The system may include one or more processors configured to acquire image data associated with multiple two-dimensional (2D) images. This image data is generated at a first time point by multiple imaging devices positioned around the patient. The system can determine the location of one or more lesions at the first time point based on the multiple 2D images. The system can determine the future location of one or more lesions at a second time point after the first time point based on the location of the one or more lesions and trajectories representing the expected motion of the one or more lesions in three-dimensional (3D) space. The system can generate control signals based on the future locations of the one or more lesions to move a linear accelerator (LINAC) from its current orientation at the first time point to its future orientation at the second time point, thereby adjusting the beam path of the linear accelerator.
[0009] In some aspects, the multiple 2D images may include one or more X-ray images. One or more processors, which can be configured to acquire image data, are configured to acquire one or more X-ray images from multiple devices. The X-ray image may include a two-dimensional (2D) representation of at least a portion of the patient at a first time point. In some aspects, one or more processors, which can be configured to determine the location of one or more lesions at the first time point, are configured to: determine the pose of one or more lesions based on the multiple 2D images. In some aspects, one or more processors may also be configured to: generate a trajectory representing the expected motion of one or more lesions based on the location of one or more previous lesions at the first time point and the future location of one or more previous lesions at a second time point. In at least one aspect, one or more processors configured to generate a trajectory representing the expected motion of one or more lesions may be configured to: provide the location of one or more lesions at the first time point to an adaptive motion model, so that the adaptive motion model generates an output representing the trajectory.
[0010] In some aspects, one or more processors configured to provide the locations of one or more lesions to an adaptive motion model can be configured to: provide the locations of one or more lesions at a first time point to the adaptive motion model. The adaptive motion model can be configured to predict a set of future locations for the one or more lesions between the first time point and a second time point. In at least one aspect, the system can obtain future location data associated with this set of future locations for the one or more lesions. The system can determine a trajectory based on this set of future locations.
[0011] In another aspect, one or more processors may also be configured to: determine a beam direction view (BEV) of one or more lesions at a second time point based on the LINAC's attitude at a second time point and the future location of one or more lesions.
[0012] In another embodiment, the method may include: acquiring image data associated with multiple two-dimensional (2D) images by one or more processors. The image data is generated by multiple imaging devices positioned around the patient at a first time point. The method may include: determining, by one or more processors, the location of one or more lesions at the first time point based on the multiple 2D images. In some aspects, the method may include: determining, by one or more processors, the future location of one or more lesions at a second time point after the first time point, based on the location of one or more lesions and trajectories representing the expected motion of one or more lesions in three-dimensional (3D) space. The method may include: generating, by one or more processors, control signals based on the future locations of one or more lesions to move a linear accelerator (LINAC) from its current pose at the first time point to its future pose at the second time point, to adjust the beam path of the linear accelerator.
[0013] In one aspect, multiple 2D images may include one or more X-ray images. Acquiring image data may include: acquiring one or more X-ray images from multiple imaging devices by one or more processors. The X-ray images may include a two-dimensional (2D) representation of at least a portion of the patient at a first time point.
[0014] In some aspects, determining the location of one or more lesions at a first time point may include: determining the pose of one or more lesions based on multiple 2D images by one or more processors. The method may also include: generating a trajectory representing the expected motion of one or more lesions by one or more processors based on the location of one or more previous lesions at the first time point and the future location of one or more previous lesions at a second time point. Generating a trajectory representing the expected motion of one or more lesions may include: providing the location of one or more lesions at the first time point to an adaptive motion model by one or more processors, causing the adaptive motion model to generate an output representing the trajectory. In some aspects, providing the location of one or more lesions to the adaptive motion model may include: providing the location of one or more lesions at the first time point to the adaptive motion model by one or more processors. The adaptive motion model may be configured to predict a set of future locations of one or more lesions between the first time point and the second time point. The method may include: obtaining future location data associated with this set of future locations of one or more lesions by one or more processors. The method includes determining a trajectory based on this set of future locations. In some aspects, the method may also include: determining a beam orientation view (BEV) of one or more lesions at a second time point by one or more processors based on the pose of the LINAC at the second time point and the future location of one or more lesions.
[0015] In another embodiment, a non-transitory computer-readable medium may store instructions thereon that, when executed by one or more processors, cause one or more processors to acquire image data associated with multiple two-dimensional (2D) images. The image data is generated at a first time point by multiple imaging devices positioned around the patient. The instructions may cause one or more processors to determine the location of one or more lesions at the first time point based on the multiple 2D images. The instructions may also cause one or more processors to determine the future location of one or more lesions at a second time point after the first time point, based on the location of the one or more lesions and trajectories representing the expected motion of the one or more lesions in three-dimensional (3D) space. Furthermore, the instructions may cause one or more processors to generate control signals based on the future locations of the one or more lesions to move a linear accelerator (LINAC) from its current pose at the first time point to its future pose at the second time point, thereby adjusting the beam path of the linear accelerator.
[0016] In one aspect, the multiple 2D images may include one or more X-ray images. Instructions causing one or more processors to acquire image data may cause one or more processors to acquire one or more X-ray images from one or more imaging devices. The X-ray images may include a two-dimensional (2D) representation of at least a portion of the patient at a first time point. In some aspects, instructions causing one or more processors to determine the location of one or more lesions at a first time point may cause one or more processors to determine the pose of one or more lesions based on the multiple 2D images. In some aspects, instructions may also cause one or more processors to generate trajectories representing the expected motion of one or more lesions based on the location of one or more previous lesions at a first time point and the future location of one or more previous lesions at a second time point.
[0017] On the other hand, instructions that cause one or more processors to generate trajectories representing the expected motion of one or more lesions can cause one or more processors to: provide the positions of one or more lesions at a first time point to an adaptive motion model, so that the adaptive motion model generates an output representing the trajectory. Instructions that cause one or more processors to provide the positions of one or more lesions to the adaptive motion model can cause one or more processors to: provide the positions of one or more lesions at a first time point to the adaptive motion model. The adaptive motion model can be configured to predict a set of future positions of one or more lesions between a first time point and a second time point. Instructions can be configured to cause one or more processors to obtain future position data associated with this set of future positions of one or more lesions. Instructions can cause one or more processors to determine a trajectory based on this set of future positions.
[0018] In one embodiment, a system is disclosed for predicting the location of a gross target volume within a patient during radiotherapy by modeling the motion of one or more lesions of a gross target volume over time. The system may include one or more processors configured to: acquire image data associated with a plurality of two-dimensional (2D) images generated by a plurality of imaging devices positioned around the patient and representing the location of one or more lesions within a first time period. The one or more processors may be configured to provide the location of one or more lesions to a sequence triangulation network to generate an estimated trajectory representing the motion of one or more lesions within the first time period. In some aspects, the one or more processors may be configured to: provide the estimated trajectory to a motion prediction network to generate a predicted trajectory representing the expected motion of one or more lesions in three-dimensional (3D) space within a second time period and determine the future location of one or more lesions based on the location of one or more lesions and the predicted trajectory. In some aspects, the one or more processors may be configured to: generate control signals based on the future location of one or more lesions to move a linear accelerator (LINAC) from a first pose to a second pose to adjust the beam path of the linear accelerator.
[0019] In some implementations, one or more processors can be configured to determine the location of one or more lesions within a first time period based on multiple 2D images. One or more processors that provide image data to a sequence triangulation network can be configured to provide image data to the sequence triangulation network, wherein the sequence triangulation network is trained based on multiple four-dimensional computed tomography (CT) scans corresponding to multiple previously observed patients. In some aspects, the sequence triangulation network can be trained based on multiple estimated 2D images extracted from four-dimensional CT scans generated for one or more patients.
[0020] In some aspects, one or more processors configured to provide the location of one or more lesions to a sequence triangulation network can be configured to adjust the sequence triangulation network based on at least a portion of the location of the one or more lesions. Multiple 2D images may include one or more X-ray images, and one or more processors configured to acquire image data can be configured to acquire one or more X-ray images from multiple imaging devices, said one or more X-ray images including a two-dimensional (2D) representation of at least a portion of the patient within a first time period. The location of one or more lesions within the first time period can represent the pose of one or more lesions in 3D space. In some aspects, one or more processors can also be configured to determine a beam orientation view (BEV) of one or more lesions based on the future location of one or more lesions and the pose of the LINAC at a second time point.
[0021] In another embodiment, a method is disclosed for predicting the location of a gross target volume within a patient during radiotherapy by modeling the motion of one or more lesions of a gross target volume over time. The method may include: acquiring image data associated with a plurality of two-dimensional (2D) images via one or more processors, the image data being generated by a plurality of imaging devices positioned around the patient and representing the location of one or more lesions within a first time period. In some aspects, the method may include: providing the location of one or more lesions to a sequence triangulation network via one or more processors, such that the sequence triangulation network generates an estimated trajectory representing the motion of one or more lesions within the first time period. In some aspects, the method may include: providing the estimated trajectory to a motion prediction network via one or more processors, such that the motion prediction network generates a predicted trajectory representing the expected motion of one or more lesions in a three-dimensional (3D) space within a second time period. In some aspects, the method may include: determining the future location of one or more lesions based on the location of one or more lesions and the predicted trajectory via one or more processors. In some aspects, the method may include: generating control signals based on the future location of one or more lesions via one or more processors to move a linear accelerator (LINAC) from a first pose to a second pose to adjust the beam path of the linear accelerator.
[0022] In some aspects, the method may further include: determining the location of one or more lesions within a first time period by one or more processors based on multiple 2D images. Providing image data to a sequence triangulation network may include: providing image data to the sequence triangulation network by one or more processors, wherein the sequence triangulation network is trained based on multiple four-dimensional computed tomography (CT) scans corresponding to multiple previously observed patients. In some aspects, the sequence triangulation network may be trained based on multiple estimated 2D images extracted from four-dimensional CT scan images generated for one or more patients. In at least some aspects, providing the location of one or more lesions to the sequence triangulation network may include: adjusting the sequence triangulation network based on at least a portion of the location of one or more lesions by one or more processors. The multiple 2D images may include one or more X-ray images, and obtaining the image data may include: obtaining one or more X-ray images from multiple imaging devices by one or more processors, said one or more X-ray images including a two-dimensional (2D) representation of at least a portion of the patient within a first time period.
[0023] In some aspects, the position of one or more lesions during a first time period can represent the pose of one or more lesions in 3D space. In some aspects, the method may also include: determining a beam orientation view (BEV) of one or more lesions by one or more processors based on the future positions of one or more lesions and the pose of the LINAC at a second time point.
[0024] In yet another embodiment, a non-transitory computer-readable medium storing instructions thereon is disclosed, which, when executed by one or more processors, cause one or more processors to: acquire image data associated with a plurality of two-dimensional (2D) images generated by a plurality of imaging devices positioned around a patient and representing the location of one or more lesions within a first time period. The instructions may cause one or more processors to: provide the location of one or more lesions to a sequence triangulation network to generate an estimated trajectory representing the motion of one or more lesions within the first time period. In some aspects, the instructions may cause one or more processors to: provide the estimated trajectory to a motion prediction network to generate a predicted trajectory representing the expected motion of one or more lesions in three-dimensional (3D) space within a second time period; and determine the future location of one or more lesions based on the location of one or more lesions and the predicted trajectory. In some aspects, the instructions may cause one or more processors to: generate control signals based on the future location of one or more lesions to move a linear accelerator (LINAC) from a first pose to a second pose to adjust the beam path of the linear accelerator.
[0025] In at least some aspects, the instructions may also cause one or more processors to determine the location of one or more lesions within a first time period based on multiple 2D images. In some aspects, the instructions causing one or more processors to provide image data to the sequence triangulation network may cause the sequence triangulation network to be trained based on multiple four-dimensional computed tomography (CT) scans corresponding to multiple previously observed patients. In some aspects, the sequence triangulation network may be trained based on multiple estimated 2D images extracted from four-dimensional CT scans generated for one or more patients. Attached Figure Description
[0026] Non-limiting embodiments of this disclosure are described by way of example and with reference to the accompanying drawings, which are schematic and not intended to be drawn to scale. Unless indicated as background art, the drawings illustrate various aspects of this disclosure.
[0027] Figure 1 A diagram is shown of a system for predicting the location of a gross target volume within a patient during radiotherapy, according to an embodiment.
[0028] Figure 2 A flowchart is shown, according to an embodiment, of a process for predicting the location of a gross target volume within a patient during radiotherapy.
[0029] Figure 3 An example implementation of a process for predicting the location of a gross target volume within a patient during radiotherapy, according to an embodiment, is shown.
[0030] Figure 4 An example of a target motion model that can be established during implementation is shown according to an embodiment.
[0031] Figure 5 An accuracy diagram of a trajectory predicted using a simulated X-ray imager according to an embodiment is shown.
[0032] Figure 6 A flowchart is shown, according to an embodiment, of a process for predicting the location of a gross target volume within a patient during radiotherapy.
[0033] Figure 7 An example implementation of a process for predicting the location of a gross target volume within a patient during radiotherapy, according to an embodiment, is shown.
[0034] Figure 8 An example of respiration-induced movement of a general target volume according to an embodiment is shown. Detailed Implementation
[0035] Reference will now be made to the illustrative embodiments depicted in the accompanying drawings, and described using specific language. Nevertheless, it should be understood that this is not intended to limit the scope of the claims or this disclosure. Changes and further modifications to the inventive features illustrated herein, as well as other applications of the subject matter principles illustrated herein (which will be apparent to those skilled in the art and are included within the scope of this disclosure), are configured to be considered within the scope of the subject matter disclosed herein. Other embodiments may be used, or other changes may be made, without departing from the spirit or scope of this disclosure. The illustrative embodiments described in the detailed description are not intended to limit the subject matter presented.
[0036] Figure 1Components of a system 100 for predicting the location of a gross target volume within a patient during radiotherapy, according to an embodiment, are shown. System 100 may include an analysis server 114a, a system database 114b, a treatment planning system 111, electronic data sources 120a-d (unless otherwise stated, each electronic data source is individually referred to as electronic data source 120 and collectively as electronic data source 120), end-user devices 140a-c (unless otherwise stated, each end-user device is individually referred to as end-user device 140 and collectively as end-user device 140), an administrator computing device 150, a medical device 160, and a medical device computer 162. Figure 1 The various components depicted herein may belong to a radiation therapy clinic where patients can receive radiation therapy (in some cases via one or more radiation therapy machines (e.g., medical device 160) located within the clinic). System 100 is not limited to the components described herein, but may include other components not shown for brevity, which are configured to be considered within the scope of the embodiments described herein.
[0037] The components mentioned above can be interconnected via network 130. Examples of network 130 may include, but are not limited to, private or public local area networks (LANs), wireless local area networks (WLANs), metropolitan area networks (MANs), wide area networks (WANs), and the Internet. Network 130 can perform wired or wireless communication according to one or more standards or via one or more transmission media. Communication via network 130 can be performed according to various communication protocols, such as Transmission Control Protocol and Internet Protocol (TCP / IP), User Datagram Protocol (UDP), and IEEE communication protocols. In one example, network 130 can perform wireless communication according to the Bluetooth specification set or other standards or proprietary wireless communication protocols. In another example, network 130 may also include communication via cellular networks, such as GSM (Global System for Mobile Communications), CDMA (Code Division Multiple Access), and EDGE (Enhanced Global Evolution Data) networks.
[0038] Analysis server 114a can be any computing device, including processors and non-transitory machine-readable storage devices capable of performing the various tasks and processes described herein. Analysis server 114a can use various processors, such as central processing units (CPUs) and graphics processing units (GPUs). Examples of such computing devices include, but are not limited to, workstation computers, laptops, server computers, etc. While system 100 includes a single analysis server 114a, analysis server 114a can include any number of computing devices operating in a distributed computing environment, such as a cloud environment.
[0039] The analysis server 114a can generate and display an electronic platform configured to use the treatment planning system 111 to receive patient information and input from users (e.g., clinicians) (such as the utility functions and updated utility functions described herein) and output the execution results of the treatment planning system 111. The electronic platform may include a graphical user interface (GUI) displayed by a display device of one or more electronic data sources 120, end-user devices 140, medical devices 160, or administrator computing devices 150. Examples of electronic platforms generated and hosted by the analysis server 114a may be web-based applications or websites configured to be displayed on various electronic devices, such as mobile devices, tablets, personal computers, etc.
[0040] For example, the information displayed by the electronic platform may include input elements for receiving data associated with the patient being treated, synchronizing one or more sensors, and displaying predictions generated by the treatment planning system 111. For example, the analysis server 114a may execute the treatment planning system 111 (e.g., a system such as a treatment planner configured or trained to generate beam configurations that can be used to configure the medical device 160 during patient treatment, as described herein). The analysis server 114a may then display the results to a clinician or directly revise one or more operational attributes of the medical device 160.
[0041] Electronic data source 120 can be any computing device, including processors and non-transitory machine-readable storage devices capable of performing the various tasks and processes described herein. For example, electronic data source 120 can represent various computing devices that contain, retrieve, or access data associated with medical device 160, such as data associated with operational information of currently or previously performed radiotherapy (e.g., electronic log files or electronic configuration files), data associated with currently or previously monitored patients (e.g., computed tomography (CT) images, magnetic resonance imaging (MRI) scans, tumor location, deformation information, etc.), or study participants. For example, analysis server 114a can use clinic computer 120a, medical professional device 120b, server 120c (associated with a clinician or clinic), and database 120d (associated with a clinician or clinic) to retrieve / receive data associated with medical device 160. Analysis server 114a can retrieve data from end-user device 140, generate datasets, and use these datasets to configure treatment planning system 111 (e.g., models implemented by treatment planning system 111). Analysis server 114a can execute various algorithms to transform raw data received / retrieved from electronic data source 120 into machine-readable objects that can be stored and processed by other analysis processes described herein.
[0042] End-user device 140 can be any computing device, including a processor and non-transitory machine-readable storage medium capable of performing the various tasks and processes described herein. Examples of end-user device 140 may be a workstation computer, laptop computer, tablet computer, or server computer (but are not limited thereto). In operation, various users as described herein, such as clinicians, can use end-user device 140 to access a graphical user interface (GUI) operatively managed by analytics server 114a, or additionally access the execution results of treatment planning system 111. Specifically, end-user device 140 may include clinic computer 140a, clinic server 140b, and medical professional device 140c. Although referred to herein as “end-user” devices, these devices are not always operated by end-users. For example, end-users cannot directly use clinic server 140b. However, results stored on clinic server 140b can be used to populate various GUIs accessed by end-users via medical professional device 140c. In some embodiments, end-user device 140 may be associated with one or more clinicians who are involved in generating one or more treatment plans for a patient (e.g., participating in the preparation of one or more treatment plans).
[0043] Administrator computing device 150 may represent a computing device operated by a system administrator. Administrator computing device 150 may be configured to display radiotherapy attributes generated by analytics server 114a (e.g., various analytical metrics determined during the training of one or more machine learning models or systems); monitor various treatment planning systems 111 utilized by analytics server 114a, electronic data source 120, or end-user device 140; review feedback; or facilitate the training or retraining (calibration) of treatment planning systems 111 maintained by analytics server 114a.
[0044] In some embodiments, medical device 160 may be a diagnostic imaging device or a therapeutic delivery device (also referred to as a radiotherapy system). For example, medical device 160 may include one or more computed tomography (CT) scanners such as cone-beam CT (CBCT) scanners, linear accelerators (LINACs) such as Varian® TrueBeam® linear accelerators, proton beam therapy systems that use accelerated protons to precisely irradiate tumors (referred to as proton beam systems), or other similar devices configured to deliver energy to a target tissue associated with the patient (referred to as a gross target volume) and, in some cases, to measure the energy delivered to protect the target tissue. Medical device 160 may also include one or more sensors configured to monitor the patient being treated. That is, medical device 160 or analysis server 114a may communicate with a variety of sensors that can monitor the patient's external biosignals. Non-limiting examples of sensors may include three-dimensional (3D) surface mechanisms, as well as optical (or other) sensors configured to monitor the patient's movement (e.g., how the patient moves or breathes). In some embodiments, medical device 160 may receive data associated with a treatment plan from (a plurality of) medical device computers 162, thereby causing medical device 160 to operate according to the treatment plan.
[0045] Treatment planning system 111 may be stored in system database 114b. Treatment planning system 111 can be trained using data received / retrieved from electronic data source 120 and can be executed using data received from end-user devices, medical devices 160, or sensors 163. In some embodiments, treatment planning system 111 may reside locally or in a clinic-specific data repository. In various embodiments, treatment planning system 111 may use one or more deep learning engines to develop treatment plans for patients receiving radiation therapy. For example, analysis server 114a may transmit patient attributes from sensor 163 and execute treatment planning system 111 accordingly. Analysis server 114a may then display the results on one or more end-user devices 140. In some embodiments, analysis server 114a may modify one or more configurations of medical device 160 based on the predictions made by treatment planning system 111.
[0046] See Figure 2 A flowchart of a process 200 for predicting the location of a gross target volume within a patient during radiotherapy, according to an embodiment, is shown. Process 200 includes operations 202-208. However, other embodiments may include additional or alternative operations, or omit one or more operations entirely. Process 200 is described as being performed by an analysis server, which can interact with... Figure 1The analysis server 114a described herein is the same as or similar. However, one or more steps of process 200 can be performed by [the server described herein]. Figure 1 The distributed computing system described herein can run on any number of computing devices to execute. For example, one or more computing devices can execute locally on... Figure 2 Some or all of the operations described in the document.
[0047] In operation 202, the analysis server can obtain image data associated with two or more two-dimensional (2D) images generated by imaging devices. For example, the analysis server can obtain image data associated with two or more 2D images generated by one or more imaging devices positioned around the patient. These imaging devices may include X-ray machines, etc. In some embodiments, the analysis server can obtain image data from multiple imaging devices, wherein the multiple imaging devices are positioned relative to the patient in three-dimensional space. In this example, the multiple imaging devices may be associated with a medical device (e.g., which is associated with...). Figure 1 Medical devices (160 similar or identical to) such as LINACs being moved to multiple control points established by a treatment plan to deliver energy to the patient are positioned relative to the patient. Although the concepts of this disclosure are discussed with regard to LINACs, it should be understood that different types of medical devices (e.g., proton beam systems) may also be considered as alternatives to or complements to LINACs.
[0048] In one example, during RT treatment, the patient can be positioned relative to a gantry of a LINAC that supports a magnetron or klystron that generates high-energy X-rays for delivery to the patient's lesion. The LINAC may include a multi-leaf collimator (MLC) positioned along the treatment beam path and configured to shape the high-energy X-rays as they are directed toward the patient's lesion. In some embodiments, multiple imaging devices can be positioned such that they target at least a portion of the patient, including the lesion targeted as part of the treatment plan. For example, during the patient's treatment, the LINAC may be moved to multiple control points (e.g., the 3D orientation of the patient, the LINAC, and the imaging devices within a 3D space) and configured to generate energy at each control point and deliver that energy to the patient. At each control point, the leaves of the LINAC's multi-leaf collimator can shape the X-rays to optimize energy delivery and consistency with the patient's lesion, while minimizing energy delivery outside the lesion (e.g., to one or more organs at risk). In some embodiments, a set of control points, MLC blade configurations, and power levels for energy delivery can be established through a treatment plan generated prior to surgery. The analytics server can be configured to perform actions to control the operation of the LINAC, as described herein, based on the treatment plan developed for the patient.
[0049] In some embodiments, the analysis server may acquire image data at a first time point. For example, the analysis server may acquire image data at two or more time points during a pre-treatment period or during a period when the patient is receiving RT treatment. During this period, the imaging devices may be configured to generate images at multiple time points. In some embodiments, the analysis server may acquire image data from multiple imaging devices, wherein the image data includes (e.g., representing) two or more X-ray images acquired by one or more imaging devices. The X-ray images may include a 2D representation of at least a portion of the patient. For example, X-ray images acquired by one or more imaging devices may represent a corresponding 2D representation of at least a portion of the patient including GTV. As described herein, the analysis server may acquire image data at two or more different time points. To allow the analysis server to acquire additional image data, one or more imaging devices may be configured to periodically or continuously generate image data and provide said image data to the analysis server.
[0050] In operation 204, the analysis server can determine the location of one or more lesions. For example, the analysis server can determine the location of one or more lesions within the patient at a first time point corresponding to the time point at which the image data was generated. In this example, the analysis server can determine the location of one or more lesions, wherein the one or more lesions are located within the patient's GTV. In some embodiments, the location of one or more lesions may represent the location and / or orientation (e.g., pose) of at least a portion of the GTV and is established relative to one or more imaging devices. The analysis server can then determine the pose of one or more lesions (e.g., included within the GTV). For example, the analysis server can determine the pose of one or more lesions in the three-dimensional space in which the patient is receiving RT treatment based on multiple 2D images acquired by imaging devices positioned relative to the patient.
[0051] In some embodiments, the analysis server can determine the location of one or more lesions by using one or more models to generate trajectories representing the expected motion of one or more lesions. For example, the analysis server can use a model (e.g., a neural network, an adaptive motion model, etc.) to generate trajectories representing the expected motion of one or more lesions. In this example, the model can be configured to receive data associated with the location of one or more lesions at a first time point and generate an output representing the trajectory. The data associated with the location of one or more lesions can be represented using multiple 2D images acquired at the first time point in time when the image data is generated. In some embodiments, the analysis server can determine the trajectory based on the model's output. For example, the analysis server can cause the model to determine (e.g., extract, etc.) the trajectory of one or more lesions over a time period starting from the time point in time when the image data is generated. The trajectory can represent the motion of one or more lesions (or at least a portion of one or more lesions) in 3D space. This motion can be represented as an offset along the X, Y, and Z axes. In some examples, this motion can also be represented by a change in velocity over a time period.
[0052] In some embodiments, the analysis server can use a model to generate trajectories, which is trained based on the locations of one or more prior lesions at a first time point and the locations of one or more prior lesions at future time points. For example, the training dataset can be built with the locations (e.g., poses) of multiple lesions at the first time point. In an example, during, for instance, an earlier RT treatment, as the imaging device tracks one or more prior lesions, the locations of one or more prior lesions can be determined based on generating 2D images similar to those described above. The training dataset can also include the corresponding locations of one or more prior lesions at a second time point. These locations at the second time point can be offset by time intervals established from previous time points.
[0053] The model can then be trained using a training dataset. For example, the analytics server can use the training dataset to train a model to output a prediction in response to the location of one or more previous lesions received during RT treatment (e.g., established from two or more 2D images acquired during RT treatment). In one example, the analytics server can train the model by providing it with the location of one or more previous lesions at a first time point, prompting the model to generate an output. In this example, the output could represent a trajectory indicating the predicted motion of one or more lesions. The analytics server can then compare the output (e.g., the predicted trajectory, which, when applied to the location of the GTV, causes the GTV to be repositioned to the predicted location) with the initially represented location of one or more previous lesions. In some examples, the analytics server can determine the future location of one or more previous lesions at a second time point (e.g., based on the motion of one or more previous lesions from their first location according to the trajectory predicted by the model) and compare the future location with a known future location established from the training dataset. The analytics server can then compare the difference between the predicted location and the known future location, determine a loss based on that difference, and update one or more weights of the model to reduce the loss when the analytics server subsequently executes the model. The analytics server can iteratively repeat this process until the model converges (e.g., the model generates a prediction that leads to a future location with a corresponding loss that satisfies a threshold of acceptable accuracy in response to the model). This can allow for consistent trajectory predictions of one or more lesions during subsequent RT treatments performed by the analytics server.
[0054] In some embodiments, once trained, the analytics server can use the model to generate trajectories representing the expected motion of one or more lesions, wherein the model is trained on a dataset representing the positions of one or more lesions as they move over time. For example, the analytics server can provide the positions of one or more lesions at a first time point as input to the model. In this example, the analytics server can prompt the model to execute based on the positions of one or more lesions at the first time point to generate output. The output can represent trajectories indicating the expected motion of one or more lesions. For example, the output can represent the trajectories of one or more lesions in 3D space, the change in acceleration of one or more lesions over time as they move in 3D space, etc.
[0055] In operation 206, the analysis server can determine the future location of one or more lesions. For example, in response to the analysis server generating one or more trajectories representing the expected movement of one or more lesions, the analysis server can determine the future location of one or more lesions. As described herein, the analysis server can determine the future location of one or more lesions within a predetermined time period starting from the time point from which the image data is generated. For example, based on a predetermined time interval after the image data is generated, the analysis server can determine the future location of one or more lesions in the 3D space where the patient is located. In this example, the time interval can be associated with (e.g., correspond to) the time period in which one or more lesions move between locations established by a training dataset. In some embodiments, the analysis server then iteratively repeats one or more of operations 202-206 and predicts the future location of one or more lesions in 3D space during RT treatment.
[0056] In some embodiments, based on the location of one or more lesions at a first time point and a trajectory generated by a model, the analysis server can determine the future location of one or more lesions. For example, by applying the trajectory to points established for one or more lesions from image data generated at the first point in the time period, the analysis server can determine the future location. In this example, the analysis server can apply a transformation to points representing one or more lesions in 3D space based on the trajectory to model the movement of one or more lesions and use the model to predict the future location of one or more lesions.
[0057] In step 208, the analysis server can generate control signals to move the medical device from its current position to a future orientation. For example, the analysis server can determine the orientation of the medical device at a first time point. At the first time point, based on the relative position of the patient to the 3D space in which the patient is located and / or the relative position of the medical device, the analysis server can configure the medical device to generate energy and transmit it toward the patient's GTV. In this example, the analysis server can cause the medical device to generate and transmit energy according to a treatment plan. For example, the analysis server can be configured to control the operation of the medical device according to a treatment plan established prior to the patient's RT treatment, wherein the treatment plan indicates one or more control points for moving the medical device between control points and one or more configurations of the medical device (e.g., one or more power levels for energy transmission, one or more blade configurations, etc.). In some embodiments, the analysis server can then predict the future position of the GTV during the RT treatment, as described above. For example, when the medical device moves from a control point toward a control point and is activated to deliver energy to the patient's GTV, the analysis server can predict the future position of the GTV. The analysis server can then compare the predicted future position with the position established by the treatment plan and control the operation of the medical device such that the medical device aims at the GTV at the (predicted) future position.
[0058] In some embodiments, the analysis server may determine a beam orientation view (BEV). For example, the analysis server may determine a beam orientation view that represents the view of the GTV relative to one or more components of a medical device (e.g., the collimator of a LINAC). This beam orientation view may be represented as a 2D representation of the GTV as seen from one or more components of the LINAC. In some embodiments, the analysis server may use the beam orientation view when aiming the GTV with the medical device during the execution of an RT treatment plan.
[0059] See Figure 3 An example implementation 300 of a process for predicting the location of a GTV (including one or more lesions thereof) within a patient during RT treatment, according to an embodiment, is shown. Figure 3 As shown, implementation 300 involves the execution of the target motion modeling network 304. In some examples, the target motion modeling network 304 can be related to the above-mentioned... Figure 2 The described models (e.g., neural networks, adaptive motion models, etc.) are the same as or similar. In some embodiments, one or more operations related to the implementation of 300 may be performed by an analysis server, which is similar to... Figure 1 Analysis server 114a and / or about Figure 2 The analysis servers discussed are the same or similar.
[0060] In some embodiments, implementation 300 may involve one or more operations performed by a target motion modeling network 304. For example, the target motion modeling network 304 may be configured to receive one or more past timestamps as input, corresponding to images 302a acquired during RT treatment by multiple imaging devices positioned around the patient. In this example, the target motion modeling network 304 may be configured to receive the location of the imaging devices used to generate image 302a, represented as the position / orientation (e.g., pose) of the imaging devices in 3D space where the patient is receiving treatment using a medical device (e.g., LINAC), as described herein. The target motion modeling network 304 can then generate a representation of a 3D target trajectory 306 (e.g., as described above regarding...). Figure 2 One or more operations are performed on the output corresponding to the described trajectory. These 3D target trajectories 306 may correspond to the expected motion of one or more lesions (e.g., GTVs) within a time period that begins at a first time point (e.g., represented by a past timestamp) and a second time point (e.g., represented by a future timestamp). In some embodiments, the 3D target trajectories 306 may be projected onto a 2D X-ray image plane using a projection matrix associated with multiple imaging devices (e.g., X-ray images, virtual imaging devices, etc., as described herein) for generating image 302a. In some embodiments, the estimated 3D locations may be projected onto the 2D X-ray image plane using the projection matrix for comparison with actual 2D observations of one or more lesions. In response to comparing the estimated 3D locations with the 2D observations, a loss may be calculated between the predicted location at the first time point established by the timestamp and the known location of one or more lesions (e.g., at a future time point) established based on the 3D target trajectories 306. During training the target motion modeling network 304, future timestamps and ground truth observations may be compared with the predicted locations of one or more lesions to iteratively adjust the parameters of the target motion modeling network 304. This allows the target motion modeling network 304 to be configured to reconstruct and predict the motion of one or more lesions within the GTV in the same implementation 300.
[0061] In some embodiments, the target motion modeling network 304 can be configured (e.g., trained) based on meta-learning of implicit functions, where the implicit functions are modeled as deep neural networks whose parameters are controlled and learned by another deep neural network. In one embodiment, the implicit neural representation can be used in conjunction with periodic activation functions to model the implicit functions. Specifically, the target motion modeling network 304 can use a set of learnable sine functions as periodic activation functions, which are more flexible in modeling semi-periodic lesion motions compared to a predetermined set of basis vectors. In some embodiments, meta-learning can involve performing a model-agnostic meta-learning (MAML) framework to enable rapid adaptation of model parameters. In the example, the MAML framework can adapt global model parameters pre-trained using a large number of offline cohort samples, allowing these parameters to quickly adapt to specific samples. In the example described herein, cohort samples from multiple patients can be used to obtain a global model (e.g., the global target motion modeling network 304), allowing for rapid fine-tuning of the global model from short-term observations of specific patients to reconstruct 3D trajectories and predict future target motions.
[0062] Figure 4 An example 400 of target motion modeling, which can be established during the execution of implementation 300 according to an embodiment, is shown. As illustrated, the ground truth “GT” trajectory and the estimated or predicted “PD” trajectory are shown. The first two columns 402 and 404 show the projected trajectories in the kV (X-ray) and MV (beam) coordinate systems, respectively. The third column 406 shows the 3D trajectories on three axes, where the values within the normalized timestamps from -1.0 to 0.5 are reconstructed, while the values from 0.5 to 1.0 are predicted. The fourth column 408 shows a magnified view of the first five predicted signals, while the fifth column 410 shows a magnified view of all predicted signals within the timestamps from 0.5 to 1.0.
[0063] Figure 5 A graph 500 showing the accuracy of the predicted trajectory using a simulated X-ray imager is presented. More specifically, a graph 500 showing the accuracy of the predicted trajectory is presented, where the simulated X-ray imager has a sampling rate of 5.2 Hz and a gantry movement speed of 5 degrees / second. It can be seen that for signal prediction within 500 milliseconds, the average error of the predicted trajectory is less than 1.0 mm; for signal prediction within 1000 milliseconds, the average error of the predicted trajectory is less than 1.5 mm.
[0064] Figure 6 A flowchart is shown, according to an embodiment, of a process for predicting the location of a gross target volume within a patient during radiotherapy.
[0065] See Figure 6A flowchart of a process 600 for predicting the location of a gross target volume within a patient during radiotherapy, according to an embodiment, is shown. Process 600 includes operations 602 to 610. However, other embodiments may include additional or alternative operations, or one or more operations may be omitted entirely. Process 600 is described as being performed by an analysis server, which can interact with… Figure 1 The analysis server 114a described herein is the same as or similar. However, one or more steps of process 200 can be performed by [the server described herein]. Figure 1 The distributed computing system described herein can run on any number of computing devices to execute. For example, one or more computing devices can execute locally on... Figure 6 Some or all of the operations described in the document.
[0066] In operation 602, the analysis server can obtain image data associated with two or more two-dimensional (2D) images generated by imaging devices. For example, the analysis server can obtain image data associated with two or more 2D images generated by multiple imaging devices positioned around the patient, similar to the above description. Figure 2 The process described in 200. Imaging devices may include X-ray machines, etc. In some embodiments, the analysis server may acquire image data from multiple imaging devices, wherein the multiple imaging devices are positioned relative to the patient in 3D space. In this example, the multiple imaging devices may be associated with a medical device such as LINAC (e.g., with...). Figure 1 The LINAC is positioned relative to the patient along with 160 identical or similar medical devices and is moved to multiple control points established through treatment planning to deliver energy to the patient. While the concepts in this disclosure are discussed with respect to LINACs, it should be understood that different types of medical devices (such as proton beam systems, etc.) may be considered as alternatives to or complements to LINACs.
[0067] In one example, during RT treatment, a patient can be positioned relative to a LINAC rack supporting a magnetron or klystron, as described herein, for generating high-energy X-rays to be delivered to the patient's lesion. During the patient's treatment, the LINAC can be moved to multiple control points and configured to generate and deliver energy to the patient at each control point. At each control point, the LINAC's MLC blades can shape the X-rays to optimize energy delivery and consistency with the patient's lesion, while minimizing energy delivery outside the lesion (e.g., one or more organs at risk). In some embodiments, a set of control points, MLC blade configurations, and power levels for energy delivery are established through a treatment plan generated preoperatively. It should be understood that an analytics server can be configured to control the operation of the LINAC by performing actions as described herein, based on the treatment plan developed for the patient.
[0068] In some embodiments, the analysis server may acquire image data after generating image data within a first time period. For example, the analysis server may acquire image data at one or more time points within the first time period, either while the patient is receiving treatment or before the patient receives treatment. In this example, the imaging device may be configured to generate images at multiple time points during the period in which the patient is being observed (e.g., as described herein, one or more models are adjusted to account for movement of one or more anatomical structures due to reasons such as the patient's breathing) and / or being treated. In some embodiments, the analysis server may acquire image data from one or more imaging devices, wherein the image data includes (e.g., representing) two or more X-ray images acquired by one or more imaging devices. The X-ray images may include a 2D representation of at least a portion of the patient. For example, the X-ray images acquired by the imaging devices may include a corresponding 2D representation of at least a portion of the patient (the GTV being treated is located in this at least a portion). To allow the analysis server to acquire additional image data, multiple imaging devices may be configured to generate image data periodically or continuously and provide the image data to the analysis server during patient observation.
[0069] In some embodiments, the analysis server can determine the location of one or more lesions represented by image data at time points within a first time period (e.g., similar to the description above regarding process 200). For example, the analysis server can compare representations of one or more events in 2D images at time points within the first time period and determine the location of one or more lesions in response to the comparison. In some examples, the location of one or more lesions can be represented as 2D coordinates and / or 3D coordinates. Alternatively or additionally, the location of one or more lesions can indicate the location and / or orientation (e.g., pose) of one or more lesions (and by extension, the GTV associated with one or more lesions). In an example, the analysis server can determine the pose of each of the one or more lesions at a time point within the first time period based on the representation of one or more lesions in a 2D image. It is understood that the analysis server can determine the location of one or more lesions, wherein the one or more lesions are located within the patient's GTV, similar to the description above regarding process 200. Figure 2 The process described in 200. Examples of techniques for determining the location of one or more lesions are also described in U.S. Patent Application No. 18 / 999,536, filed December 23, 2024 (titled “SYSTEMS AND METHODS FOR DETERMINING A LOCATION OF A GROSS TARGET VOLUME OF APATIENT”), the contents of which are incorporated herein by reference in their entirety for all purposes.
[0070] In operation 604, the analysis server can provide the locations of one or more lesions to a sequence triangulation network to generate an estimated trajectory. For example, the analysis server can provide the locations of one or more lesions to a sequence triangulation network (e.g., a model such as a neural network, deep neural network, transformer, or transformer-based network configured to perform one or more operations described herein to model one or more implicit functions) built from two or more 2D images of the patient. The sequence triangulation network can then be configured to generate an output associated with (e.g., representing the estimated trajectory) the locations of one or more lesions built from the two or more 2D images. This estimated trajectory can represent the motion of one or more lesions observed in 3D space within a time period determined based on 2D images generated before or during the patient's RT treatment.
[0071] In some embodiments, the analysis server can enable a sequence triangulation network to generate estimated trajectories of one or more lesions in a patient. For example, during RT treatment, the analysis server can acquire two or more 2D images of one or more lesions in the patient. The analysis server can then provide these two or more 2D images as input to the sequence triangulation network, causing the network to generate an output representing the estimated trajectory. This estimated trajectory can represent the relative motion of one or more lesions in the 3D space in which the patient is located.
[0072] Sequence triangulation networks can be trained using a training dataset. For example, an analytics server can train a sequence triangulation network using the training dataset to generate an estimated trajectory of the patient's lesions in response to acquiring (e.g., receiving, deriving, etc.) a 2D image of a patient (e.g., an X-ray image) and determining the location of one or more lesions (e.g., established from two or more 2D images acquired during RT treatment). In the example, the analytics server can train the sequence triangulation network by providing the location of one or more lesions, causing the sequence triangulation network to generate a corresponding output. In the example, the corresponding output can represent an estimated trajectory indicating the movement of one or more lesions observed in the patient's 2D image. The analytics server can then compare the output (e.g., the estimated trajectory) with a known trajectory established using the training dataset. The analytics server can then determine the difference between the estimated trajectory and the known trajectory (e.g., the offset between the estimated trajectory and the known trajectory), determine a loss based on this difference, and update one or more weights of the sequence triangulation network to reduce the loss when the analytics server subsequently executes the sequence triangulation network. The analytics server can iteratively repeat this process until the sequence triangulation network converges (e.g., the model generates a prediction that results in a future location with a corresponding loss that satisfies a threshold of acceptable accuracy in response to the sequence triangulation network).
[0073] In some embodiments, the known trajectory can be determined based on one or more four-dimensional CT (4DCT) scans. For example, the training dataset can be built based on 4DCT scans acquired for one or more previously observed patients. In this example, the analysis server can generate multiple 2D images in different gantry poses from the 4DCT scans of previously observed patients and include these 2D images in the training dataset. Similarly, the analysis server can generate multiple trajectories from 4DCT scans. In some embodiments, the analysis server can then use the training dataset built based on the 4DCT scans to train the sequence triangulation network as described above.
[0074] In some embodiments, during RT treatment, the analysis server can tune the sequence triangulation network. For example, the analysis server can tune the sequence triangulation network based on the location of one or more lesions acquired by an imaging device for a specific patient receiving treatment. The analysis server can determine the location of one or more lesions based on 2D images of the patient acquired before or during RT treatment and feed these locations to the sequence triangulation network to generate estimated trajectories. By providing the location of one or more lesions of the patient before or during treatment, the analysis server can train (e.g., update) the sequence triangulation network to model the motion of a specific lesion within that patient. This helps to more accurately determine the location of one or more lesions for each patient.
[0075] In operation 606, the analysis server may provide the estimated trajectory to the motion prediction network to generate a predicted trajectory. For example, the analysis server may provide the estimated trajectories of one or more lesions to the motion prediction network (e.g., a model configured to perform one or more operations described herein to model one or more implicit functions, such as a neural network, deep neural network, transformer, or transformer-based network, etc.) based on (e.g., in response to) the generation of the estimated trajectory by a sequence triangulation network. The motion prediction network may be configured to generate an output associated with (e.g., representing) the predicted trajectory of the expected motion of one or more lesions within the 3D space in which the patient is located during a second time period. Specifically, the analysis server may cause the motion prediction network to generate the predicted trajectory over a period of time following the time period represented by the estimated trajectory. The analysis server can then use the predicted trajectory to determine the location (e.g., pose) of one or more lesions, as described herein.
[0076] Motion prediction networks can be trained using a training dataset. For example, an analytics server can use the training dataset to train a motion prediction network to generate predicted trajectories for patient lesions in response to obtaining (e.g., receiving, deriving, etc.) estimated trajectories of one or more lesions. In this example, the analytics server can train the motion prediction network by providing the location and / or estimated trajectory of one or more lesions of one or more patients represented in the training dataset at a certain time point (e.g., corresponding to the final time point represented by the estimated trajectory) to generate a corresponding output. In this example, the corresponding output could represent a predicted trajectory indicating the expected motion of one or more lesions of the corresponding patient within the 3D space in which the patient is located in a second time period (after the time period established by the estimated trajectory). The analytics server can then compare the output (e.g., the predicted trajectory) with a known trajectory established from the training dataset. In response to this comparison, the analytics server can determine the difference between the predicted trajectory and the known trajectory (e.g., the offset between the predicted trajectory and the known trajectory), calculate a loss based on this difference, and update one or more weights of the motion prediction network to reduce the loss when the analytics server subsequently executes the motion prediction network. The analysis server can iteratively repeat this process until the motion prediction network converges (e.g., the model generates a prediction that results in a trajectory with a corresponding loss that satisfies a threshold of acceptable accuracy in response to the motion prediction network).
[0077] In operation 608, the analysis server can determine the future location of one or more lesions. For example, the analysis server can determine the future location of one or more lesions in response to generating a predicted motion trajectory representing one or more lesions. In this example, the analysis server can determine the future location of one or more lesions based on the location of one or more lesions represented by image data at the start of the second time period and the predicted trajectory. In some embodiments, based on the predicted trajectory, the analysis server can determine the future location (e.g., future pose) of one or more lesions based on applying a transformation to the location of one or more lesions at the start of the second time period. It should be understood that the analysis server then iteratively repeats one or more of operations 602-606 and predicts the future location of one or more lesions in 3D space during RT treatment.
[0078] In operation 610, the analysis server can generate control signals to move the medical device from its current position to a future pose. For example, the analysis server can determine the pose of the medical device at a first time point (as described above, at the start of the second time period). At the first time point, based on the relative position of the patient relative to the 3D space in which the patient is located and / or the relative position of the medical device, the analysis server can configure the medical device to generate energy and deliver it to the patient's GTV. In this example, the analysis server can cause the medical device to generate and deliver energy according to a treatment plan similar to that described in process 200. In some embodiments, as described above, the analysis server can then predict the future position of the GTV during the RT treatment process. For example, when the medical device moves from a control point to a control point and is activated to deliver energy to the patient's GTV, the analysis server can predict the future position of the GTV. The analysis server can then compare the predicted future position with the position established by the treatment plan and control the operation of the medical device such that the medical device aims at the GTV at the (predicted) future position.
[0079] In some embodiments, the analysis server may determine a beam direction view. For example, the analysis server may determine a beam direction view that represents the view of the GTV relative to one or more components of a medical device (e.g., the collimator of a LINAC). The beam direction view may be represented as a 2D representation of the GTV from the perspective of one or more components of the LINAC. In some embodiments, the analysis server may use the beam direction view when aiming the GTV with the medical device during the execution of an RT treatment plan.
[0080] Figure 7 An example implementation 700 of a process for predicting the location of a gross target volume within a patient during radiotherapy, according to an embodiment, is shown. Figure 7 As shown, implementation 700 involves executing a sequence triangulation network 702 and a motion prediction network 704. In some examples, the sequence triangulation network 702 and / or the motion prediction network 704 can be related to the above description. Figure 6 The described model is the same as or similar. In some embodiments, one or more operations described with respect to implementation 700 may be performed by an analysis server, which is similar to... Figure 1 Analysis server 114a and / or about Figure 2 The analysis servers discussed are the same or similar.
[0081] In some embodiments, implementation 700 may include generating multiple 2D projected trajectories using a 2D trajectory generator 702c. For example, one or more 4DCT images 702b may be obtained and used to determine (e.g., derive) 3D trajectories. The 2D trajectory generator 702c can then determine individual angles and timestamps from the 3D trajectories to build a training dataset and allow training of a sequence triangulation network 702 (similar to that described above) to configure the sequence triangulation network 702 to reconstruct 3D trajectories (referred to as estimated trajectories) from a sequence of 2D X-ray images (also referred to as 2D observations) from different angles within a first time period.
[0082] Based on 2D images generated by an imaging device directed at the patient during RT treatment, the trained sequence triangulation network 702' can be further updated for a specific patient. The estimated trajectory can be provided as input to the motion prediction network 704 to predict 3D trajectories over future timeframes. Both the sequence triangulation network 702' and the motion prediction network 704 can be modeled using the implicit functions described herein. In some embodiments, alternatively, an external substitution signal can be used as input to one or more sequence triangulation networks 702, 702', or the motion prediction network 704.
[0083] Figure 8 An example of movement of a gross target volume (GTV) 800 according to an embodiment is shown, the movement being caused by respiration. In some embodiments, the GTV 800 may be the same as or similar to the GTV described above. In a first state 800a, the GTV 800 may remain stationary; and during respiration 800b, the GTV 800 may move from a first position to a second position. As illustrated, whether the GTV 800 is moving or not, the GTV may be at least partially surrounded by a clinical target volume (CTV). Similarly, the CTV may also be at least partially surrounded by a planned target volume (PTV).
[0084] When moving (e.g., from a first position to a second position during an individual's respiration), the GTV 800 and CTV 801 can move in coordination with the PTV. For example, during respiration, the GTV 800, representing one or more lesions, may be significantly displaced due to movement of surrounding tissues and organs. This movement (particularly relative to the PTV) can be modeled using the techniques described herein. In the example, the CTV 801 includes not only the GTV 800 but also any possible subclinical lesions, while the internal target volume (ITV) is based on the CTV to address uncertainties caused by organ movement, and the PTV may include additional margins to account for uncertainties in treatment delivery, such as patient positioning and setup errors. As shown in 800b, during expiration, the GTV 800 and CTV 801 can move upward relative to the ITV and PTV; while during inspiration, the GTV 800 and CTV 801 can move downward relative to the ITV and PTV. It should be understood that although the movement is shown in one dimension, the GTV 800 can move in multiple directions within three-dimensional space.
[0085] The following enumerated examples will provide a better understanding of the currently disclosed technology:
[0086] 1. A system for predicting the location of a gross target volume within a patient during radiotherapy by modeling the motion of one or more lesions of a gross target volume over time, the system comprising: one or more processors configured to: acquire image data associated with a plurality of two-dimensional (2D) images generated by one or more imaging devices positioned around the patient at a first time point; determine the location of one or more lesions at the first time point based on the plurality of 2D images; determine the future location of one or more lesions at a second time point after the first time point based on the location of one or more lesions and trajectories representing the expected motion of one or more lesions in three-dimensional (3D) space; and generate control signals based on the future locations of one or more lesions to move a linear accelerator (LINAC) from a current orientation at the first time point to a future orientation at the second time point to adjust the beam path of the linear accelerator.
[0087] 2. The system according to Clause 1, wherein the plurality of 2D images include a plurality of X-ray images, and wherein one or more processors configured to acquire image data are configured to: acquire the plurality of X-ray images from one or more imaging devices, the X-ray images including a two-dimensional (2D) representation of at least a portion of the patient at a first time point. 3. The system according to any one of the foregoing examples, wherein one or more processors configured to determine the location of one or more lesions at a first time point are configured to: determine the pose of one or more lesions based on multiple 2D images. 4. The system according to any one of the foregoing examples, wherein one or more processors are further configured to: generate a trajectory representing the expected movement of one or more lesions based on the location of one or more prior lesions at a first time point and the future location of one or more prior lesions at a second time point. 5. The system according to any one of the foregoing examples, wherein one or more processors configured to generate trajectories representing the expected motion of one or more lesions are configured to: provide the position of one or more lesions at a first time point to an adaptive motion model so that the adaptive motion model generates an output representing the trajectory.
[0088] 6. The system according to any one of the foregoing examples, wherein one or more processors configured to provide the location of one or more lesions to an adaptive motion model are configured to: provide the location of one or more lesions at a first time point to the adaptive motion model, wherein the adaptive motion model is configured to predict a set of future locations of one or more lesions between the first time point and a second time point; obtain future location data associated with the set of future locations of one or more lesions; and determine a trajectory based on the set of future locations. 7. The system according to any one of the foregoing examples, wherein one or more processors are further configured to: determine a beam direction view (BEV) of one or more lesions at a second time point based on the LINAC's attitude at a second time point and the future location of one or more lesions.
[0089] 8. A method comprising: acquiring image data associated with a plurality of two-dimensional (2D) images by one or more processors, the image data being generated by one or more imaging devices positioned around a patient at a first time point; determining, by one or more processors, the location of one or more lesions at the first time point based on the plurality of 2D images; determining, by one or more processors, the future location of the one or more lesions at a second time point after the first time point based on the location of the one or more lesions and trajectories representing the expected motion of the one or more lesions in three-dimensional (3D) space; and generating, by one or more processors, control signals based on the future locations of the one or more lesions to move a linear accelerator (LINAC) from a current pose at the first time point to a future pose at the second time point, thereby adjusting the beam path of the linear accelerator. 9. The method according to any one of the foregoing examples, wherein the plurality of 2D images comprises a plurality of X-ray images, and wherein obtaining image data comprises: obtaining the plurality of X-ray images from one or more imaging devices by one or more processors, the X-ray images comprising a two-dimensional (2D) representation of at least a portion of the patient at a first time point. 10. The method according to any one of the foregoing examples, wherein determining the location of one or more lesions at a first time point comprises: determining the pose of one or more lesions by one or more processors based on multiple 2D images.
[0090] 11. The method according to any one of the foregoing examples, further comprising: generating, by one or more processors, a trajectory representing the expected movement of one or more lesions based on the location of one or more prior lesions at a first time point and the future location of one or more prior lesions at a second time point. 12. The method according to any one of the foregoing examples, wherein generating a trajectory representing the expected motion of one or more lesions comprises: providing the positions of one or more lesions at a first time point to an adaptive motion model by one or more processors, such that the adaptive motion model generates an output representing the trajectory.
[0091] 13. The method according to any one of the foregoing examples, wherein providing the location of one or more lesions to the adaptive motion model comprises: providing the location of one or more lesions at a first time point to the adaptive motion model by one or more processors, wherein the adaptive motion model is configured to predict a set of future locations of one or more lesions between the first time point and a second time point; obtaining future location data associated with the set of future locations of one or more lesions by one or more processors; and determining a trajectory based on the set of future locations.
[0092] 14. The method according to any one of the foregoing examples further includes: determining a beam direction view (BEV) of one or more lesions at a second time point by one or more processors based on the LINAC pose at a second time point and the future location of one or more lesions.
[0093] 15. A computer program storing instructions thereon, which, when executed by one or more processors, cause the one or more processors to perform the following steps: acquiring image data associated with a plurality of two-dimensional (2D) images, the image data being generated by one or more imaging devices positioned around a patient at a first time point; determining the location of one or more lesions at the first time point based on the plurality of 2D images; determining the future location of one or more lesions at a second time point after the first time point based on the location of the one or more lesions and trajectories representing the expected motion of the one or more lesions in three-dimensional (3D) space; and generating control signals based on the future locations of the one or more lesions to move a linear accelerator (LINAC) from its current pose at the first time point to its future pose at the second time point, thereby adjusting the beam path of the linear accelerator. 16. The computer program according to any one of the foregoing examples, wherein the plurality of 2D images comprises a plurality of X-ray images, and wherein instructions for causing one or more processors to acquire image data cause one or more processors to perform the steps of: acquiring a plurality of X-ray images from one or more imaging devices, the X-ray images comprising a two-dimensional (2D) representation of at least a portion of a patient at a first time point. 17. A computer program according to any of the foregoing examples, wherein instructions for causing one or more processors to determine the location of one or more lesions at a first time point cause one or more processors to perform the step of determining the pose of one or more lesions based on a plurality of 2D images. 18. The computer program according to any one of the foregoing examples, wherein the instructions further cause one or more processors to perform the step of: generating a trajectory representing the expected movement of one or more lesions based on the location of one or more prior lesions at a first time point and the future location of one or more prior lesions at a second time point. 19. A computer program according to any one of the foregoing examples, wherein instructions to cause one or more processors to generate trajectories representing the expected motion of one or more lesions cause one or more processors to perform the step of: providing the position of one or more lesions at a first time point to an adaptive motion model so that the adaptive motion model generates an output representing the trajectory.
[0094] 20. A computer program according to any one of the foregoing examples, wherein instructions for causing one or more processors to provide the location of one or more lesions to an adaptive motion model cause the one or more processors to perform the steps of: providing the location of one or more lesions at a first time point to the adaptive motion model, wherein the adaptive motion model is configured to predict a set of future locations of one or more lesions between the first time point and a second time point; obtaining future location data associated with the set of future locations of one or more lesions; and determining a trajectory based on the set of future locations.
[0095] The various illustrative logic blocks, modules, circuits, and algorithm steps described in conjunction with the embodiments disclosed herein can be implemented as electronic hardware, computer software, or a combination of both. To clearly illustrate this interchangeability between hardware and software, the various illustrative components, modules, circuits, and steps have been generally described above according to their functionality. Whether this functionality is implemented as hardware or software depends on the specific application and the design constraints imposed on the overall system. Those skilled in the art can implement the described functionality in different ways for each specific application, but such implementation choices should not be construed as departing from the scope of this disclosure or the claims.
[0096] Implementations using computer software (e.g., computer programs, computer program products, etc.) can be implemented as software, firmware, middleware, microcode, hardware description languages, or any combination thereof. Code segments or machine-executable instructions can represent programs, functions, subroutines, routines, subroutines, modules, software packages, classes, or any combination of instructions, data structures, or program statements. Code segments can be coupled to other code segments or hardware circuit systems by passing or receiving information, data, parameters, or memory contents. Information, actual parameters, formal parameters, data, etc., can be passed, forwarded, or transmitted via any suitable means, including memory sharing, message passing, token passing, network transmission, etc.
[0097] The actual software code or dedicated control hardware used to implement these systems and methods does not limit the claimed features or this disclosure. Therefore, since the operation and behavior of the system and method are described without mentioning specific software code, it should be understood that software and control hardware can be designed to implement the system and method based on the description herein.
[0098] When implemented in software, these functions can be stored as one or more instructions or code on a non-transitory computer-readable or processor-readable storage medium. The steps of the methods or algorithms disclosed herein can be embodied in a processor-executable software module, which can reside on a computer-readable or processor-readable storage medium. Non-transitory computer-readable or processor-readable media include computer storage media and tangible storage media that facilitate the transfer of computer programs from one place to another. A non-transitory processor-readable storage medium can be any medium accessible to a computer. For example, but not limited to, such a non-transitory processor-readable medium can include RAM, ROM, EEPROM, CD-ROM or other optical disc storage, disk storage or other magnetic storage devices, or any other tangible storage medium that can be used to store desired program code in the form of instructions or data structures and is accessible to a computer or processor. As used herein, “disk” and “optical disc” include compact discs (CDs), laser discs, optical discs, digital versatile discs (DVDs), floppy disks, and Blu-ray discs, where disks typically copy data magnetically, while optical discs copy data optically using lasers. Combinations of the various media described above should also be included within the scope of computer-readable media. In addition, the operation of a method or algorithm may be located on a non-transitory processor-readable medium or computer-readable medium in the form of a single or arbitrary combination or set of code or instructions, and such medium may be incorporated into a computer program product.
[0099] The foregoing description of the disclosed embodiments is intended to enable those skilled in the art to make or use the embodiments described herein and variations thereof. Various modifications to these embodiments will readily be understood by those skilled in the art, and the principles defined herein can be applied to other embodiments without departing from the spirit or scope of the subject matter disclosed herein. Therefore, this disclosure is not intended to be limited to the embodiments shown herein, but should be applied most broadly within the scope of the following claims and the principles and novel features disclosed herein.
[0100] While various aspects and embodiments have been disclosed, other aspects and embodiments are contemplated. The disclosed aspects and embodiments are for illustrative purposes only and are not intended to be limiting; the true scope and spirit are indicated by the following claims.
Claims
1. A system for predicting the location of a gross target volume within a patient during radiotherapy by modeling the motion of one or more lesions of a gross target volume over time, the system comprising: One or more processors, said one or more processors being configured to: Image data associated with multiple two-dimensional (2D) images, the image data being generated at a first time point by one or more imaging devices positioned around the patient; Based on the multiple 2D images, determine the location of one or more lesions at the first time point; Based on the location of the one or more lesions and the trajectory representing the expected movement of the one or more lesions in three-dimensional (3D) space, determine the future location of the one or more lesions at a second time point after the first time point; Based on the future position of the one or more lesions, a control signal is generated to move the linear accelerator (LINAC) from its current attitude at the first time point to its future attitude at the second time point, so as to adjust the beam path of the linear accelerator.
2. The system of claim 1, wherein the plurality of 2D images comprises a plurality of X-ray images, and The one or more processors configured to acquire the image data are configured to: Multiple X-ray images are obtained from the one or more imaging devices, the X-ray images including a two-dimensional (2D) representation of at least a portion of the patient at the first time point.
3. The system of claim 1, wherein the one or more processors configured to determine the location of the one or more lesions at the first time point are configured to: The pose of the one or more lesions is determined based on the plurality of 2D images.
4. The system of claim 3, wherein the one or more processors are further configured to: Based on the location of one or more previous lesions at a first time point and the future location of the one or more previous lesions at a second time point, the trajectory representing the expected movement of the one or more lesions is generated.
5. The system of claim 4, wherein the one or more processors configured to generate the trajectory representing the expected movement of the one or more lesions are configured to: The location of the one or more lesions at the first time point is provided to the adaptive motion model so that the adaptive motion model generates an output representing the trajectory.
6. The system of claim 5, wherein the one or more processors configured to provide the location of the one or more lesions to the adaptive motion model are configured to: The location of the one or more lesions at the first time point is provided to the adaptive motion model, wherein the adaptive motion model is configured to predict a set of future locations of the one or more lesions between the first time point and the second time point; Obtain future location data associated with the set of future locations of the one or more lesions; as well as The trajectory is determined based on the set of future locations.
7. The system of claim 1, wherein the one or more processors are further configured to: Based on the orientation of the LINAC at the second time point and the future position of the one or more lesions, a beam direction view (BEV) of the one or more lesions at the second time point is determined.
8. A method, the method comprising: Image data associated with multiple two-dimensional (2D) images is obtained by one or more processors, the image data being generated at a first time point by one or more imaging devices positioned around the patient; The one or more processors determine the location of the one or more lesions at the first time point based on the plurality of 2D images; The one or more processors determine the future position of the one or more lesions at a second time point after the first time point, based on the location of the one or more lesions and a trajectory representing the expected movement of the one or more lesions in three-dimensional (3D) space. The one or more processors generate control signals based on the future positions of the one or more lesions to move the linear accelerator (LINAC) from its current pose at the first time point to its future pose at the second time point, thereby adjusting the beam path of the linear accelerator.
9. The method of claim 8, wherein the plurality of 2D images comprises a plurality of X-ray images, and The acquisition of the image data includes: The processor acquires multiple X-ray images from the imaging device, the X-ray images including a two-dimensional (2D) representation of at least a portion of the patient at the first time point.
10. The method of claim 8, wherein determining the location of the one or more lesions at the first time point comprises: The orientation of the one or more lesions is determined by the one or more processors based on the plurality of 2D images.
11. The method of claim 10, further comprising: The one or more processors generate the trajectory representing the expected movement of the one or more lesions, based on the location of the one or more previous lesions at a first time point and the future location of the one or more previous lesions at a second time point.
12. The method of claim 11, wherein generating the trajectory representing the expected movement of the one or more lesions comprises: The one or more processors provide the location of the one or more lesions at the first time point to the adaptive motion model, so that the adaptive motion model generates an output representing the trajectory.
13. The method of claim 12, wherein providing the location of the one or more lesions to the adaptive motion model comprises: The one or more processors provide the location of the one or more lesions at the first time point to the adaptive motion model, wherein the adaptive motion model is configured to predict a set of future locations of the one or more lesions between the first time point and the second time point; The one or more processors obtain future location data associated with the set of future locations of the one or more lesions; as well as The trajectory is determined based on the set of future locations.
14. The method of claim 8, further comprising: The one or more processors determine the beam orientation view (BEV) of the one or more lesions at the second time point based on the attitude of the LINAC at the second time point and the future position of the one or more lesions.
15. A non-transitory computer-readable medium having instructions stored thereon, the instructions causing the one or more processors, when executed, to: Image data associated with multiple two-dimensional (2D) images, the image data being generated at a first time point by one or more imaging devices positioned around the patient; Based on the multiple 2D images, determine the location of one or more lesions at the first time point; Based on the location of the one or more lesions and the trajectory representing the expected movement of the one or more lesions in three-dimensional (3D) space, determine the future location of the one or more lesions at a second time point after the first time point; Based on the future position of the one or more lesions, a control signal is generated to move the linear accelerator (LINAC) from its current attitude at the first time point to its future attitude at the second time point, so as to adjust the beam path of the linear accelerator.
16. The non-transitory computer-readable medium of claim 15, wherein: The plurality of 2D images includes a plurality of X-ray images, and The instruction that causes the one or more processors to obtain the image data causes the one or more processors to: Multiple X-ray images are obtained from the one or more imaging devices, the X-ray images including a two-dimensional (2D) representation of at least a portion of the patient at the first time point.
17. The non-transitory computer-readable medium of claim 15, wherein the instruction causing the one or more processors to determine the location of the one or more lesions at the first time point causes the one or more processors to: The pose of the one or more lesions is determined based on the plurality of 2D images.
18. The non-transitory computer-readable medium of claim 17, wherein the instructions further cause the one or more processors to: Based on the location of one or more previous lesions at a first time point and the future location of the one or more previous lesions at a second time point, the trajectory representing the expected movement of the one or more lesions is generated.
19. The non-transitory computer-readable medium of claim 18, wherein the instructions causing the one or more processors to generate the trajectory representing the expected movement of the one or more lesions cause the one or more processors to: The location of the one or more lesions at the first time point is provided to the adaptive motion model so that the adaptive motion model generates an output representing the trajectory.
20. The non-transitory computer-readable medium of claim 19, wherein the instruction causing the one or more processors to provide the location of the one or more lesions to the adaptive motion model causes the one or more processors to: The location of the one or more lesions at the first time point is provided to the adaptive motion model, wherein the adaptive motion model is configured to predict a set of future locations of the one or more lesions between the first time point and the second time point; Obtain future location data associated with the set of future locations of the one or more lesions; as well as The trajectory is determined based on the set of future locations.