Image processing apparatus, radiation imaging system, image processing method, and storage medium

The image processing device addresses the issue of pixel value conversion by applying a unit conversion process, enabling accurate image generation and calculations using trained models.

JP2026017180APending Publication Date: 2026-02-04CANON KK
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
JP2024117888
Authority / Receiving Office
JP · JP
Patent Type
Applications
Current Assignee / Owner
Filing Date
2024-07-23
Publication Date
2026-02-04

AI Technical Summary

Technical Problem

Existing image processing methods using trained models fail to convert pixel values into physical lengths, leading to inappropriate calculations and desired image generation.

Method used

An image processing device that includes a processing unit to apply a unit conversion process, converting pixel values from a trained model into physical lengths.

Benefits of technology

Enables accurate conversion of pixel values into physical lengths, ensuring appropriate calculations and desired image generation.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 2026017180000001_ABST
    Figure 2026017180000001_ABST
Patent Text Reader

Abstract

To provide an image processing device capable of converting a pixel value of an image output from a learned model into a physical length.SOLUTION: An image processing apparatus comprising: an obtaining unit configured to obtain a radiation image; and a processing unit configured to obtain an energy subtraction image by inputting the obtained radiation image to a trained model, wherein the processing unit applies, to the obtained energy subtraction image, unit conversion processing of converting a pixel value into a physical length.SELECTED DRAWING: Figure 20
Need to check novelty before this filing date? Find Prior Art

Description

[Technical Field]

[0001] The present invention relates to an image processing device, a radiation imaging system, an image processing method, and a program. [Background technology]

[0002] A commonly known imaging technique is time-multiplexed spectral imaging. In time-multiplexed spectral imaging, multiple radiation beams with different average energies are irradiated onto the subject over a short period of time, and the constituent materials of the subject are distinguished by measuring the proportion of radiation beams with each average energy that pass through the subject and reach the radiation measurement surface. Specifically, energy subtraction (ES) processing of acquired radiation images with different energies (pairs of high- and low-energy images in the case of two energies) can be used to obtain energy subtraction images such as bone images and soft tissue images.

[0003] Such time-division spectral imaging is also used to generate medical radiological images, as in the technology described in Patent Document 1. Patent Document 2 proposes a method for improving the image quality of energy subtraction images acquired by these methods using an image generation model based on machine learning. [Prior art documents] [Patent documents]

[0004] [Patent Document 1] Japanese Patent Application Publication No. 2019-162358 [Patent Document 2] Japanese Patent Publication No. 2023-120851 Summary of the Invention [Problem to be solved by the invention]

[0005] When acquiring an ES image using a trained model, such as the technology described in Patent Document 2, a method may be used in which a desired post-processing is performed on an image output from the trained model to generate a target ES image. Here, in the post-processing, for example, when a calculation is performed using pixel values ​​of the output image of the trained model, it may be desirable for the pixel values ​​to be values ​​that correspond to physical units of length. However, if the pixel values ​​of the image output from the trained model are not values ​​that correspond to physical units of length, the desired image may not be generated due to the pixel values.

[0006] Therefore, one object of one embodiment of the present disclosure is to provide an image processing device that can convert pixel values ​​of an image output from a trained model into physical lengths. [Means for solving the problem]

[0007] An image processing device according to one embodiment of the present disclosure includes an acquisition unit that acquires a radiographic image, and a processing unit that acquires an energy subtraction image by inputting the acquired radiographic image into a trained model, and the processing unit applies a unit conversion process to the acquired energy subtraction image to convert pixel values ​​into physical lengths. [Effects of the Invention]

[0008] According to one embodiment of the present disclosure, pixel values ​​of an image output from a trained model can be converted into physical lengths. [Brief explanation of the drawings]

[0009] [Figure 1] 1 shows an example of the overall configuration of a radiation imaging system according to a first embodiment. [Figure 2] FIG. 2 is an equivalent circuit diagram of an example of a pixel of the radiation imaging device. [Figure 3] 10 is a timing chart showing an example of a radiation imaging operation. [Figure 4] 10 is a timing chart showing an example of a radiation imaging operation. [Figure 5] FIG. 10 is a block diagram of a correction process included in the preprocessing. [Figure 6] FIG. 1 is a block diagram of signal processing for ES processing. [Figure 7] FIG. 10 is a block diagram of a process for generating a total thickness image. [Figure 8] FIG. 10 is a block diagram of a process for generating a virtual monochromatic X-ray image. [Figure 9] 10 is a flowchart illustrating an example of a learning process. [Figure 10] FIG. 10 is a block diagram of a learning process. [Figure 11] 1 shows an example of the configuration of a machine learning model. [Figure 12] 3 shows an example of learning data according to the first embodiment. [Figure 13] 10 is a flowchart illustrating an example of an image generation process using a trained model. [Figure 14] FIG. 1 is a block diagram of an inference process using a trained model. [Figure 15] 10 shows examples of input and output images during learning and inference. [Figure 16] FIG. 10 is a block diagram showing an example of a process for generating an ES image from a radiation image. [Figure 17] FIG. 10 is a block diagram showing another example of the process of generating an ES image from a radiation image. [Figure 18] FIG. 10 is a block diagram showing another example of the process of generating an ES image from a radiation image. [Figure 19] 4 is a flowchart showing an example of a series of processes according to the first embodiment. [Figure 20] FIG. 2 is a block diagram showing an example of a series of processes according to the first embodiment. [Figure 21] FIG. 10 is a diagram for explaining a total thickness image. [Figure 22] 10 is a flowchart illustrating an example of a process for determining a scale value α. [Figure 23] FIG. 10 is a block diagram illustrating an example of a process for determining a scale value α. [Figure 24]10 is a flowchart showing an example of a series of processes according to Modification 2. [Figure 25] FIG. 10 is a block diagram showing an example of a series of processes according to Modification 2. [Figure 26] FIG. 10 is a block diagram showing another example of a series of processes according to the second modification. [Figure 27] FIG. 10 is a block diagram illustrating an example of processing using a clipped image. DETAILED DESCRIPTION OF THE INVENTION

[0010] Hereinafter, preferred embodiments of the present disclosure will be described in detail with reference to the drawings. However, the dimensions, materials, shapes, and relative positions of components described in the following embodiments are arbitrary and can be changed depending on the configuration of an apparatus to which the present disclosure is applied or various conditions. In addition, the same reference numerals are used in the drawings to indicate identical or functionally similar elements.

[0011] In this disclosure, radiation includes α-rays, β-rays, γ-rays, and the like, which are beams created by particles (including photons) emitted by radioactive decay, as well as beams with the same or higher energy levels, such as X-rays, particle beams, and cosmic rays.

[0012] Energy subtraction processing (ES processing) refers to processing that calculates differences using images (energy images) related to different radiation energies to obtain, for example, material decomposition images of bone and soft tissue, contrast agent and water, etc., information on effective atomic number and surface density, etc. The energy subtraction processing may include, for example, correction processing such as offset correction processing as pre-processing, and image processing such as contrast adjustment processing and total thickness image generation processing as post-processing.

[0013] The term "energy subtraction image (ES image)" can include, for example, a material decomposition image obtained using ES processing, an image showing effective atomic number and surface density, etc. Furthermore, the term "ES image" can also include a material decomposition image obtained using an approximation formula for ES processing described below, a radiation image of any energy, etc. Furthermore, the term "ES image" can also include an image inferred using a trained model obtained by training an image obtained using the above-mentioned ES processing, a material decomposition image obtained using three-dimensional data from CT (Computed Tomography), etc. Furthermore, hereinafter, the term "ES image" can also include a total thickness image generated by adding together material decomposition images such as bone images and soft tissue images.

[0014] In the following, a machine learning model refers to a learning model based on a machine learning algorithm. Specific examples of machine learning algorithms include nearest neighbor algorithms, naive Bayes algorithms, decision trees, and support vector machines. Another example is deep learning, which uses a neural network to generate features and connection weighting coefficients for learning. Any of the above algorithms that can be used can be used as appropriate and applied to the following embodiments. Furthermore, training data refers to training data, and is composed of a pair of input data and output data. Furthermore, supervised answer data refers to output data from training data (training data).

[0015] Furthermore, a trained model refers to a machine learning model that follows any machine learning algorithm, such as deep learning, and that has been trained (learned) in advance using appropriate training data (learning data). However, although a trained model is obtained in advance using appropriate training data, it does not mean that no further learning is performed, and additional learning can also be performed. Additional learning can also be performed after the device is installed at the site of use. Note that obtaining output data from input data using a trained model is sometimes referred to as inference.

[0016] (First embodiment) As described above, when an image is output from a trained model, the pixel values ​​of the output image may not correspond to physical units of length. For example, when the output of the trained model is normalized between 0 and 1, more specifically, when the pixel values ​​of an image used as output data of training data are normalized between 0 and 1, the pixel values ​​of the image output from the trained model may not correspond to physical units of length. Furthermore, for example, due to an error in inference by the trained model, the pixel values ​​of the image output from the trained model may not correspond to physical units of length.

[0017] On the other hand, in post-processing of processing using a trained model, calculations such as arbitrary image conversion and image analysis are performed using pixel values ​​of an image output from the trained model, and a desired image or analysis result is obtained using the calculation results. However, if a calculation is performed using pixel values ​​that do not conform to physical units of length, an appropriate calculation result may not be obtained, and the desired image or analysis result may not be obtained.

[0018] Therefore, the image processing device according to this embodiment performs a unit conversion process for converting pixel values ​​of an image output from a trained model into physical lengths (values ​​corresponding to units of physical lengths). By performing this unit conversion process, the image processing device according to this embodiment can obtain appropriate calculation results using images obtained using the trained model, and can acquire desired images and analysis results. Hereinafter, the radiation imaging system, image processing device, and image processing method according to this embodiment will be described with reference to FIGS. 1 to 23.

[0019] (Configuration of Radiation Imaging System) 1 is a block diagram showing an example of the overall configuration of a radiation imaging system according to this embodiment. The radiation imaging system according to this embodiment includes a radiation generating device 101, a radiation control device 102, a control device 103, a radiation imaging device 104, an input unit 150, and a display unit 120.

[0020] The radiation generating device 101 includes a radiation source such as a tube. The radiation generating device 101 generates radiation under the control of the radiation control device 102. The radiation control device 102 includes a control circuit, a processor, etc. Based on the control of the control device 103, the radiation control device 102 controls the radiation generating device 101 so that radiation is irradiated toward the subject Su and the radiation imaging device 104. More specifically, the radiation control device 102 can control imaging conditions such as the irradiation angle of radiation emitted by the radiation generating device 101, the radiation focal position, the tube voltage, and the tube current.

[0021] The radiation control device 102, the radiation imaging device 104, the input unit 150, and the display unit 120 are connected to the control device 103, and the control device 103 can control these. The control device 103 can perform, for example, various controls related to radiation imaging and image processing for generating ES images.

[0022] The control device 103 is provided with an acquisition unit 131, a generation unit 132, a processing unit 133, a display control unit 134, and a storage unit 135. The acquisition unit 131 can acquire images captured by the radiation imaging device 104, images generated by the generation unit 132, etc. The acquisition unit 131 can also acquire various images from an external device (not shown) connected to the control device 103 via a network such as the Internet.

[0023] The generation unit 132 can generate a radiation image from an image (image information) captured using the radiation imaging device 104 and acquired by the acquisition unit 131. The generation unit 132 can generate energy images (high energy image and low energy image) related to radiation energy from an image captured using the radiation imaging device 104 by irradiating radiation of any energy, for example. A method for generating an energy image will be described later.

[0024] The processing unit 133 generates an ES image based on an energy image (radiographic image) using the trained model. The trained model and a method for generating an ES image will be described later. The processing unit 133 can also perform image processing and analysis using the generated ES image, etc., as well as post-processing such as unit conversion processing on an image output from the trained model. Furthermore, the processing unit 133 can function as an example of a learning unit that trains a machine learning model.

[0025] The display control unit 134 controls the display of the display unit 120 and can display, for example, information on the subject Su (examination object), information on radiography, various acquired and generated images, etc. on the display unit 120. The storage unit 135 can store, for example, information on the subject Su, information on radiography, and various acquired and generated images. The storage unit 135 can also store programs and the like for the control device 103 to perform various processes.

[0026] Here, the control device 103 can be configured by a computer provided with a processor and a memory. The control device 103 may be configured by a general computer or a computer dedicated to the radiation control system. The control device 103 may be, for example, a personal computer, and a desktop PC, a notebook PC, a tablet PC (portable information terminal), or the like may be used. Furthermore, the control device 103 may be configured as a cloud-based computer in which some of the components are located in an external device.

[0027] Furthermore, each component of the control device 103 other than the storage unit 135 may be configured by a software module executed by a processor such as a CPU (Central Processing Unit) or an MPU (Micro Processing Unit). The processor may be, for example, a GPU (Graphical Processing Unit) or an FPGA (Field-Programmable Gate Array). Each component may be configured by a circuit that performs a specific function, such as an ASIC. The storage unit 135 may be configured by any storage medium, such as an optical disk such as a hard disk, or a memory.

[0028] The display unit 120 is configured with any monitor, and displays various information such as information about the subject Su, various images, and a mouse cursor in accordance with the operation of the input unit 150, under the control of the display control unit 134. The input unit 150 is an input device that issues instructions to the control device 103, and specifically includes a keyboard and a mouse. Note that the display unit 120 may be configured with a touch panel display, and in this case, the display unit 120 can also be used as the input unit 150.

[0029] The radiation imaging device 104 detects radiation that is irradiated from the radiation generation device 101 and passes through the subject Su, and captures a radiation image. The radiation imaging device 104 can be configured as, for example, an FPD. The radiation imaging device 104 is provided with a phosphor 141 that converts radiation into visible light, and a two-dimensional detector 142 that detects the visible light. The two-dimensional detector 142 includes a sensor in which pixels 20 that detect visible light are arranged in an array of X columns and Y rows, and outputs image information according to the detected radiation dose.

[0030] An example of the pixel 20 will now be described with reference to Fig. 2. Fig. 2 shows an equivalent circuit diagram of an example of the pixel 20. The pixel 20 is provided with a photoelectric conversion element 201 and an output circuit unit 202. The photoelectric conversion element 201 may typically include a photodiode. The output circuit unit 202 is provided with an amplifier circuit unit 204, a clamp circuit unit 206, a sample-and-hold circuit unit 207, and a selection circuit unit 208.

[0031] The photoelectric conversion element 201 includes a charge accumulation section, which is connected to the gate of a MOS transistor 204a in the amplifier circuit section 204. The source of the MOS transistor 204a is connected to a current source 204c via a MOS transistor 204b. The MOS transistor 204a and the current source 204c form a source follower circuit. The MOS transistor 204b is an enable switch that turns on when an enable signal EN supplied to its gate becomes active level, putting the source follower circuit into an operating state.

[0032] In the example shown in FIG. 2, the charge storage section of the photoelectric conversion element 201 and the gate of the MOS transistor 204a form a common node. This node functions as a charge-voltage converter that converts the charge stored in the charge storage section of the photoelectric conversion element 201 into a voltage. A voltage V (=Q / C) appears in the charge-voltage converter, which is determined by the charge Q stored in the charge storage section and the capacitance C of the charge-voltage converter. The charge-voltage converter is connected to a reset potential Vres via a reset switch 203. When a reset signal PRES becomes active level, the reset switch 203 is turned on, and the potential of the charge-voltage converter is reset to the reset potential Vres.

[0033] The clamp circuit unit 206 clamps the noise output by the amplifier circuit unit 204 in response to the reset potential of the charge-voltage converter using a clamp capacitor 206a. In other words, the clamp circuit unit 206 is a circuit for canceling the noise described above from the signal output from the source follower circuit in response to the charge generated by photoelectric conversion in the photoelectric conversion element 201. This noise includes kTC noise generated during reset. Clamping is achieved by first setting the clamp signal PCL to an active level to turn on the MOS transistor 206b, and then setting the clamp signal PCL to an inactive level to turn off the MOS transistor 206b. The output side of the clamp capacitor 206a is connected to the gate of the MOS transistor 206c. The source of the MOS transistor 206c is connected to a current source 206e via a MOS transistor 206d. The MOS transistor 206c and the current source 206e form a source follower circuit. The MOS transistor 206d is an enable switch that turns on when an enable signal EN0 supplied to its gate becomes active, activating the source follower circuit.

[0034] A signal output from the clamp circuit unit 206 in response to charges generated by photoelectric conversion of the photoelectric conversion element 201 is written as an optical signal to the capacitor 207Sb via the switch 207Sa when the optical signal sampling signal TS becomes active. The signal output from the clamp circuit unit 206 when the MOS transistor 206b is turned on immediately after resetting the potential of the charge-voltage converter is a clamp voltage. The clamp voltage is written as noise to the capacitor 207Nb via the switch 207Na when the noise sampling signal TN becomes active. This noise includes an offset component of the clamp circuit unit 206. The switch 207Sa and the capacitor 207Sb form a signal sample-and-hold circuit 207S, and the switch 207Na and the capacitor 207Nb form a noise sample-and-hold circuit 207N. The sample-and-hold circuit unit 207 includes the signal sample-and-hold circuit 207S and the noise sample-and-hold circuit 207N.

[0035] When the driving circuit unit drives the row selection signal to an active level, the signal (optical signal) held in the capacitor 207Sb is output to the signal line 21S via the MOS transistor 208Sa and the row selection switch 208Sb. At the same time, the signal (noise) held in the capacitor 207Nb is output to the signal line 21N via the MOS transistor 208Na and the row selection switch 208Nb. The MOS transistor 208Sa forms a source follower circuit together with a constant current source (not shown) provided on the signal line 21S. Similarly, the MOS transistor 208Na forms a source follower circuit together with a constant current source (not shown) provided on the signal line 21N. The MOS transistor 208Sa and the row selection switch 208Sb form the signal selection circuit unit 208S, and the MOS transistor 208Na and the row selection switch 208Nb form the noise selection circuit unit 208N. The selection circuit section 208 includes a signal selection circuit section 208S and a noise selection circuit section 208N.

[0036] The pixel 20 may have an addition switch 209S that adds optical signals from multiple adjacent pixels 20. In addition mode, an addition mode signal ADD becomes active, and the addition switch 209S is turned on. As a result, the capacitances 207Sb of adjacent pixels 20 are connected to each other by the addition switch 209S, and the optical signals are averaged. Similarly, the pixel 20 may have an addition switch 209N that adds noises from multiple adjacent pixels 20. When the addition switch 209N is turned on, the capacitances 207Nb of adjacent pixels 20 are connected to each other by the addition switch 209N, and the noises are averaged. The adder 209 includes an addition switch 209S and an addition switch 209N.

[0037] The pixel 20 may have a sensitivity change unit 205 for changing the sensitivity. The pixel 20 may include, for example, a first sensitivity change switch 205a, a second sensitivity change switch 205'a, and associated circuit elements. When the first change signal WIDE becomes active, the first sensitivity change switch 205a is turned on, and the capacitance value of the first additional capacitor 205b is added to the capacitance value of the charge-voltage conversion unit. This reduces the sensitivity of the pixel 20. When the second change signal WIDE2 becomes active, the second sensitivity change switch 205'a is turned on, and the capacitance value of the second additional capacitor 205'b is added to the capacitance value of the charge-voltage conversion unit. This further reduces the sensitivity of the pixel 20. By adding a function to reduce the sensitivity of the pixel 20 in this way, it becomes possible to receive a larger amount of light and widen the dynamic range. When the first change signal WIDE becomes active, the enable signal ENw may be turned on to cause the MOS transistor 204'a to operate as a source follower instead of the MOS transistor 204a.

[0038] The radiation imaging device 104 reads the output of the pixel circuit as described above and converts it into digital values ​​(image information) using an AD converter (not shown). The radiation imaging device 104 transfers the image information converted into digital values ​​to the control device 103. This allows the acquisition unit 131 of the control device 103 to acquire the radiographed image.

[0039] (Radiation imaging operation for ES processing) Next, before describing the processing according to this embodiment, the radiation imaging operation for performing the ES processing and the ES processing will be described with reference to FIGS.

[0040] Figures 3 and 4 show examples of various drive timings in radiography operations for performing ES processing in a radiographic imaging system. Figure 3 shows an example of radiography operations when a relatively inexpensive tube that does not allow switching of tube voltage (energy) is used, while Figure 4 shows an example of radiography operations when a tube that does allow switching of tube voltage is used. The waveforms in the figures, with time on the horizontal axis, show the timing of X-ray exposure, synchronization signal, resetting of photoelectric conversion element 201, driving of sample-and-hold circuit unit 207, and readout of image from signal line 21. Note that in Figures 3 and 4, the waveform for X-rays represents tube voltage. Also, while black and white areas are provided for X-rays, this is simply done to make it easier to distinguish the timing.

[0041] First, the example shown in FIG. 3 will be described. In this example, X-rays are emitted after the photoelectric conversion element 201 is reset. Ideally, the X-ray tube voltage is a rectangular wave, but it takes a finite amount of time for the tube voltage to rise and fall. In particular, when the exposure time is short with pulsed X-rays, the tube voltage can no longer be considered a rectangular wave, and takes on a waveform like that shown in the X-ray waveform in FIG. 3. For this reason, the energy of the X-rays differs during the rise, stable, and fall periods of the X-rays.

[0042] Therefore, after X-rays 301 in the rising phase are irradiated, sampling is performed by the noise sample and hold circuit 207N, and further, after X-rays 302 in the stable phase are irradiated, sampling is performed by the signal sample and hold circuit 207S. Thereafter, the difference between the signal on signal line 21N and the signal on signal line 21S is read out as an image. At this time, the signal (R1) of X-rays 301 in the rising phase is held in the noise sample and hold circuit 207N, and the sum of the signal (R1) of X-rays 301 in the rising phase and the signal (B) of X-rays 302 in the stable phase is held in the signal sample and hold circuit 207S. Therefore, an image 304 corresponding to the signal (B) of X-rays 302 in the stable phase is read out as the difference between the signal on signal line 21N and the signal on signal line 21S.

[0043] Next, after the irradiation of X-rays 303 in the falling phase and the readout of image 304 are completed, sampling is again performed by signal sample-and-hold circuit 207S. Thereafter, photoelectric conversion element 201 is reset, and sampling is again performed by noise sample-and-hold circuit 207N, and the difference between the signal on signal line 21N and the signal on signal line 21S is read out as an image. At this time, the noise sample-and-hold circuit 207N holds a signal in a state where no X-rays are irradiated. Furthermore, the signal sample-and-hold circuit 207S holds the sum of the signal (R1) of X-rays 301 in the rising phase, the signal (B) of X-rays 302 in the stable phase, and the signal (R2) of X-rays 303 in the falling phase. Therefore, image 306, which corresponds to the sum of the signal (R1) of X-rays 301 in the rising phase, the signal (B) of X-rays 302 in the stable phase, and the signal (R2) of X-rays 303 in the falling phase, is read out as the difference between the signal on signal line 21N and the signal on signal line 21S. Thereafter, by calculating the difference between image 306 and image 304, image 305 corresponding to the sum of the signal (R1) of X-ray 301 in the rising phase and the signal (R2) of X-ray 303 in the falling phase is obtained.

[0044] The timing for resetting the sample-and-hold circuit unit 207 and the photoelectric conversion element 201 is determined using a synchronization signal 307 indicating that X-ray exposure has started from the radiation generation device 101. A method for detecting the start of radiation exposure may include measuring the tube current of the radiation generation device 101 and determining whether the current value exceeds a predetermined threshold. Alternatively, a configuration may be used in which, after resetting of the photoelectric conversion element 201 is completed, signals from the pixels 20 are repeatedly read out and whether the pixel values ​​exceed a predetermined threshold is determined. Furthermore, a configuration may be used in which an X-ray detector other than the two-dimensional detector 142 is built into the radiation imaging device 104 and whether the measured value exceeds a predetermined threshold is determined. In either case, sampling by the signal sample-and-hold circuit 207S, sampling by the noise sample-and-hold circuit 207N, and resetting of the photoelectric conversion element 201 are performed after a predetermined time has elapsed since the input of the synchronization signal 307.

[0045] In this way, it is possible to obtain an image 304 corresponding to the stable period of the pulsed X-rays and an image 305 corresponding to the sum of the rising and falling periods of the pulsed X-rays. Because the energy of the X-rays irradiated when generating these two images is different, ES processing can be performed by performing calculations between the two images.

[0046] Next, an example of radiation imaging operation when using a tube with switchable tube voltage will be described with reference to Fig. 4. This example differs from the example shown in Fig. 3 in that the X-ray tube voltage is actively switched.

[0047] In this example, first, the photoelectric conversion element 201 is reset, and then low-energy X-rays 401 are irradiated. Then, sampling is performed by the noise sample-and-hold circuit 207N, and the tube voltage is switched to irradiate high-energy X-rays 402. After the high-energy X-rays 402 are irradiated, sampling is performed by the signal sample-and-hold circuit 207S. Then, the tube voltage is switched to irradiate low-energy X-rays 403. Furthermore, the difference between the signals on the signal line 21N and the signal line 21S is read out as an image. At this time, the noise sample-and-hold circuit 207N holds the signal (R1) of the low-energy X-rays 401, and the signal sample-and-hold circuit 207S holds the sum of the signal (R1) of the low-energy X-rays 401 and the signal (B) of the high-energy X-rays 402. Therefore, an image 404 corresponding to the signal (B) of the high-energy X-ray 402 is read out as the difference between the signal on the signal line 21N and the signal on the signal line 21S.

[0048] Next, after the irradiation of low-energy X-rays 403 and the readout of image 404 are completed, sampling is again performed by signal sample-and-hold circuit 207S. Thereafter, photoelectric conversion element 201 is reset, and sampling is again performed by noise sample-and-hold circuit 207N, and the difference between the signal on signal line 21N and the signal on signal line 21S is read out as an image. At this time, the noise sample-and-hold circuit 207N holds a signal in a state where no X-rays are irradiated. Furthermore, the signal sample-and-hold circuit 207S holds the sum of the signal (R1) of low-energy X-rays 401, the signal (B) of high-energy X-rays 402, and the signal (R2) of low-energy X-rays 403. Therefore, an image 406 corresponding to the sum of the signal (R1) of low-energy X-rays 401, the signal (B) of high-energy X-rays 402, and the signal (R2) of low-energy X-rays 403 is read out as the difference between the signal on signal line 21N and the signal on signal line 21S. Thereafter, by calculating the difference between image 406 and image 404, image 405 corresponding to the sum of the signal (R1) of low-energy X-ray 401 and the signal (R2) of low-energy X-ray 403 is obtained.

[0049] The synchronization signal 407 is the same as in the example shown in Fig. 3. In this way, by acquiring images while actively switching the tube voltage, it is possible to make the energy difference between low-energy images and high-energy images larger than in the method of Fig. 3. The method of acquiring low-energy images and high-energy images is not limited to this. For example, an FPD including a stacked sensor formed of two types of materials (phosphors) with different X-ray absorption rates may be used as the radiation imaging device 104, and low-energy and high-energy images may be acquired based on the outputs from the sensors of each layer obtained by a single exposure.

[0050] (ES processing) Next, the ES processing method will be described. In addition to the signal processing of the ES processing, the ES processing can include correction processing as pre-processing and image processing as post-processing.

[0051] First, the correction process, which is pre-processing, will be described with reference to Fig. 5. Fig. 5 is a block diagram of the correction process. Note that in the following, this embodiment will be described assuming that a radiation imaging operation is performed in accordance with the example shown in Fig. 3.

[0052] 3 without irradiating the radiation imaging device 104 with X-rays, and the acquisition unit 131 acquires the captured images. At this time, two images corresponding to images 304 and 306 are read out, with the first image (image 304) being image F_Odd and the second image (image 306) being image F_Even. Images F_Odd and F_Even are images corresponding to fixed pattern noise (FPN) of the radiation imaging device 104.

[0053] Next, in a state where there is no subject, the radiation imaging device 104 is irradiated with X-rays to perform imaging using the driving method shown in Fig. 3, and the acquisition unit 131 acquires the captured images. At this time, two images corresponding to images 304 and 306 are read out, with the first image (image 304) being image W_Odd and the second image (image 306) being image W_Even. Images W_Odd and W_Even are images corresponding to the sum of signals due to the FPN of the radiation imaging device 104 and X-rays.

[0054] Therefore, by subtracting image F_Odd from image W_Odd and image F_Even from image W_Even, images WF_Odd and WF_Even can be obtained from which the FPN of the radiation imaging device 104 has been removed. In this embodiment, such correction processing is called offset correction.

[0055] Image WF_Odd is an image corresponding to X-rays 302 in the stable period, and image WF_Even is an image corresponding to the sum of X-rays 301 in the rising period, X-rays 302 in the stable period, and X-rays 303 in the falling period. Therefore, by subtracting image WF_Odd from image WF_Even, an image corresponding to the sum of X-rays 301 in the rising period and X-rays 303 in the falling period is obtained. The energies of X-rays 301 in the rising period and X-rays 303 in the falling period are lower than the energy of X-rays 302 in the stable period. Therefore, by subtracting image WF_Odd from image WF_Even, a low-energy image W_Low in the absence of an object is obtained. Furthermore, a high-energy image W_High in the absence of an object is obtained from image WF_Odd. In this embodiment, this correction process is called color correction.

[0056] Next, with a subject present, the radiation imaging device 104 is irradiated with X-rays to perform imaging using the driving method shown in Fig. 3, and the acquisition unit 131 acquires the captured images. At this time, two images corresponding to images 304 and 306 are read out, with the first image (image 304) designated as image X_Odd and the second image (image 306) designated as image X_Even. The generation unit 132 performs offset correction and color correction on these images in the same manner as when there is no subject, thereby obtaining a low-energy image X_Low when there is a subject and a high-energy image X_High when there is a subject.

[0057] Here, if the thickness of the subject is d, the linear attenuation coefficient of the subject is μ, the output of pixel 20 when there is no subject is I0, and the output of pixel 20 when there is a subject is I, the following equation (1) holds.

number

number

[0058] Therefore, by dividing the low-energy image X_Low when there is an object by the low-energy image W_Low when there is no object, the image of the attenuation rate at low energy (low-energy image Im L ) is obtained. Similarly, by dividing the high-energy image X_High when there is an object by the high-energy image W_High when there is no object, an image of the attenuation rate H at high energy (high-energy image Im H Hereinafter, such a correction process will be referred to as gain correction. In the radiation imaging system, the generation unit 132 performs correction processes including the offset correction, color correction, and gain correction as described above, thereby obtaining a low-energy image Im L and high-energy image Im H can be generated and obtained.

[0059] Next, the signal processing of the ES processing will be described with reference to Fig. 6. Fig. 6 is a block diagram of the signal processing of the ES processing. In the signal processing of the ES processing, the low-energy image Im obtained by the correction processing described with reference to Fig. L and high energy images H From the bone thickness image (bone image Im B ) and soft tissue thickness image (soft tissue image Im S ) is found.

[0060] First, let E be the energy of the X-ray photon, N(E) be the number of photons at energy E, B be the thickness of the bone, and S be the thickness of the soft tissue. Let μ be the linear attenuation coefficient of the bone at energy E. B (E), the linear attenuation coefficient of soft tissue at energy E is μ S (E), if the attenuation ratio is I / I0, the following equation (3) holds.

number

[0061] The number of photons N(E) at energy E is the X-ray spectrum. The X-ray spectrum can be obtained by simulation or actual measurement for the tube voltage. Also, the linear attenuation coefficient μ of bone at energy E isB (E) and the linear attenuation coefficient μ of soft tissue at energy E S (E) can be obtained from databases such as those of the National Institute of Standards and Technology (NIST). Therefore, by using equation (3), it is possible to calculate the attenuation ratio I / I0 for any bone thickness B, soft tissue thickness S, and X-ray spectrum N(E).

[0062] Here, the spectrum of low energy X-rays is N L (E) High-energy X-ray spectrum H If (E), the following equation (4) holds.

number

[0063] Here, we will explain the case where the Newton-Raphson method, which is a type of iterative solution, is used as a typical method for solving nonlinear simultaneous equations. First, let us define the number of iterations of the Newton-Raphson method as m, and the bone thickness after the mth iteration as B. m , the thickness of the soft tissue after the mth iteration is S m Then, the attenuation rate of high energy after the mth iteration, H m , and the low-energy attenuation rate L after the mth iteration m is expressed by the following equation (5).

number

number

[0064] At this time, the bone thickness B after the m+1th iteration m+1 and soft tissue thickness S m+1 is expressed by the following equation (7) using the attenuation rate of high energy H and the attenuation rate of low energy L.

number

number

number

[0065] By repeating this calculation, the high-energy attenuation rate H m The difference between the measured high-energy attenuation rate H and the measured high-energy attenuation rate L approaches 0. The same is true for the low-energy attenuation rate L. As a result, the bone thickness B after the mth iteration m converges to the bone thickness B, and the m-th soft tissue thickness S m converges to the thickness S of the soft tissue. In this way, the nonlinear simultaneous equations shown in equation (4) can be solved. Therefore, by calculating equation (4) for all pixels, the low-energy image Im L and high energy images H From the bone image B and soft tissue imaging S can be obtained.

[0066] In this example, for the sake of simplicity, the bone thickness B and the soft tissue thickness S are calculated by ES processing, but this embodiment is not limited to this. For example, the thickness of water and the thickness of a contrast agent may be calculated by ES processing. In this case, the linear attenuation coefficient of water at energy E and the linear attenuation coefficient of a contrast agent at energy E may also be obtained from a database such as NIST. The ES processing makes it possible to calculate the thickness of any two types of material. Furthermore, an image of the effective atomic number Z and an image of the surface density D may be obtained from an image of the attenuation rate L at low energy and an image of the attenuation rate H at high energy obtained by the correction shown in FIG. 5. The effective atomic number Z is the equivalent atomic number of the mixture, and the surface density D is the density [g / cm] of the subject. 3 ] and the thickness of the subject [cm].

[0067] In the above description, the nonlinear simultaneous equations are solved using the Newton-Raphson method. However, the method for solving the nonlinear simultaneous equations is not limited to this. For example, iterative methods such as the least squares method and bisection method may also be used. Furthermore, the method for calculating the bone thickness B and the soft tissue thickness S is not limited to solving the nonlinear simultaneous equations using an iterative method. For example, the bone thickness B and the soft tissue thickness S for various combinations of high-energy attenuation rate H and low-energy attenuation rate L may be calculated in advance to generate a table, and the bone thickness B and the soft tissue thickness S may be calculated quickly by referring to the table.

[0068] Next, image processing for generating a total thickness image and a virtual monochromatic X-ray image in post-processing of such ES processing will be described. First, the processing for generating a total thickness image will be described with reference to FIG. 7. FIG. 7 is a block diagram of image processing for generating a total thickness image. In the image processing shown in FIG. 7, the image of bone thickness (bone image Im B ) and soft tissue thickness image (soft tissue image Im S ) and the sum of these images is the total thickness image Im T can be obtained.

[0069] In the human body, bones are surrounded by soft tissue, so the thickness of the soft tissue in the area where the bone is located is reduced by the thickness of the bone. S When gradation processing is applied to the image, the soft tissue image Im S On the other hand, bone contrast appears in the soft tissue image Im S and bone image Im B Adding , backfills the loss of bone thickness in the soft tissue image. Therefore, the total thickness image Im T is an image in which the contrast of bones has been removed. T A time-direction filter such as a recursive filter or a space-direction filter such as a Gaussian filter may be applied to the image to reduce noise, and then the image may be subjected to gradation processing and displayed.

[0070] Next, image processing for generating a virtual monochromatic X-ray image will be described with reference to Fig. 8. Fig. 8 is a block diagram of image processing for generating a virtual monochromatic X-ray image. In the image processing shown in Fig. 8, the bone image Im obtained by the signal processing shown in Fig. 6 is B and soft tissue imaging Im S Here, a virtual monochromatic X-ray image is an image that is assumed to be obtained when X-rays of a single energy are irradiated.

[0071] Virtual monochromatic X-ray images are used in Dual Energy CT, which combines ES processing and three-dimensional reconstruction. Virtual monochromatic X-ray images can suppress beam hardening artifacts and metal artifacts. For example, when the energy of virtual monochromatic X-rays is set to E V Then, the virtual monochromatic X-ray image V is calculated by the following equation (10).

number

[0072] In addition, the energy E of the virtual monochromatic X-ray VBy changing the linear attenuation coefficient μ of bone, the CNR (contrast to noise ratio) of the virtual monochromatic X-ray image can be improved. B (E) is the linear attenuation coefficient μ of soft tissue S (E). However, the energy E V The larger the difference between the linear attenuation coefficient μB(E) of bone and the linear attenuation coefficient μS(E) of soft tissue becomes smaller. Therefore, the energy E of the virtual monochromatic X-ray V By setting a large value, the increase in noise in the virtual monochromatic X-ray image due to noise in the bone image is suppressed. On the other hand, the virtual monochromatic X-ray energy E V The smaller the linear attenuation coefficient μ of bone, B (E) and the linear attenuation coefficient μ of soft tissue S The contrast of the virtual monochromatic X-ray image increases because the difference in energy E V By setting an appropriate value, the CNR of the virtual monochromatic X-ray image can be improved, and for example, the amount of contrast agent used in radiography can be reduced.

[0073] In this embodiment, a virtual monochromatic X-ray image is generated from the bone thickness B and the soft tissue thickness S, but the present disclosure is not limited to this form. After calculating the effective atomic number Z and the surface density D by ES processing, a virtual monochromatic X-ray image may be generated using the effective atomic number Z and the surface density D. In addition, when a plurality of energies E V A composite X-ray image may be generated by combining multiple virtual monochromatic X-ray images generated in step 1. A composite X-ray image is an image that is expected to be obtained when irradiating an object with X-rays of a given spectrum.

[0074] (Learning process) In the above description, bone images and soft tissue images are generated by ES processing using high-energy images and low-energy images. In contrast, in this embodiment, an ES image is generated based on a radiological image, which is an input image, using a machine learning model, and an ES image of a different type from the ES image is generated based on the ES image and the input image using an approximation formula for ES processing. For ease of explanation, an example of generating a total thickness image using a machine learning model will be described below. Here, the learning process of the machine learning model according to this embodiment will be described with reference to FIGS. 9 to 12. FIG. 9 is a flowchart showing an example of the learning process of the machine learning model according to this embodiment.

[0075] In step S901, training data is prepared. Here, radiation images such as high-energy images or low-energy images are used as input data for the training data, and total thickness images are used as output data for the training data. Note that either high-energy images or low-energy images may be used as input data for the training data. In the following, an example will be described in which high-energy images are used as input data and total thickness images are used as output data.

[0076] As shown in FIGS. 1 to 6, the training data can be prepared by acquiring captured high-energy images and low-energy images, performing ES processing to acquire bone images and soft-tissue images, and generating a total thickness image by adding the bone images and soft-tissue images. More specifically, the acquisition unit 131 acquires the captured high-energy images and low-energy images. Then, the processing unit 133 performs ES processing on the high-energy images and low-energy images to generate bone images and soft-tissue images. The processing unit 133 also adds the generated bone images and soft-tissue images to generate a total thickness image. This allows the control device 103 to acquire radiographic images to be used as input data for the training data and a total thickness image to be used as output data for the training data.

[0077] Alternatively, a virtual monochromatic X-ray image may be generated from the bone image and the soft tissue image, the virtual monochromatic X-ray image may be used as input data for the learning data, and the total thickness image may be used as output data for the learning data. In this case, the processing unit 133 may generate the virtual monochromatic X-ray image from the bone image and the soft tissue image.

[0078] In the above description, a configuration has been assumed in which high-energy images and low-energy images are acquired and ES processing is performed to obtain bone images and soft tissue images, but the present disclosure is not limited to such a configuration.

[0079] The acquisition unit 131 may acquire high-energy images, low-energy images, bone images, soft-tissue images, total thickness images, and virtual monochromatic X-ray images from an external device such as a server (not shown) connected to the control device 103. Furthermore, the processing unit 133 may generate bone images and soft-tissue images from the high-energy images and low-energy images acquired from the external device by the acquisition unit 131. Similarly, the processing unit 133 may generate total thickness images and virtual monochromatic X-ray images from the bone images and soft-tissue images acquired from the external device by the acquisition unit 131.

[0080] Alternatively, images for training data may be prepared by obtaining two-dimensional images by forward projection from three-dimensional data acquired by CT. For example, after calculating the ratio of bone to soft tissue from the CT value, two-dimensional bone images and soft tissue images can be obtained separately by forward projection.

[0081] As an example, the acquisition unit 131 can acquire three-dimensional data obtained by CT from an external device. Then, the processing unit 133 acquires three-dimensional data of bone and three-dimensional data of soft tissue based on the ratio of bone to soft tissue calculated from the CT value for the acquired three-dimensional data. Furthermore, the processing unit 133 can perform a known forward projection process on these to acquire two-dimensional bone images and soft tissue images. Then, the processing unit 133 can generate a total thickness image based on the acquired bone images and soft tissue images. The acquisition unit 131 may also acquire bone images and soft tissue images based on the three-dimensional data obtained by CT from an external device. Note that these processes are merely examples, and the control device 103 may acquire bone images and soft tissue images based on the three-dimensional data obtained by CT using any known process.

[0082] In this case, a virtual monochromatic X-ray image can be generated from the bone image and soft tissue image obtained by forward projection using the method shown in Fig. 8, and the pair of the virtual monochromatic X-ray image and the total thickness image can be used as learning data. In this case as well, the processing unit 133 can generate a virtual monochromatic X-ray image from the bone image and soft tissue image obtained by forward projection.

[0083] Here, an example will be described in which a total thickness image is used as output data of the training data, but the configuration of the training data is not limited to this. ES images, such as material decomposition images of bone images and soft tissue images, can also be used as output data of the training data. Therefore, the training data may be, for example, a pair of a high-energy image and a bone image, a pair of a low-energy image and a bone image, or a pair of a virtual monochromatic X-ray image and a bone image. Similarly, the training data may be, for example, a pair of a high-energy image and a soft-tissue image, a pair of a low-energy image and a soft-tissue image, or a pair of a virtual monochromatic X-ray image and a soft-tissue image.

[0084] Next, in step S902, the processing unit 133 performs preprocessing of the training data. Machine learning requires a huge amount of data. If the number of data pieces in the training data prepared in step S901 is insufficient, the processing unit 133 can increase the training data by data expansion. Methods of data expansion that can be used include clipping, which extracts a portion of an image, and rotating, scaling, and flipping an image.

[0085] Alternatively, the processing unit 133 may perform data expansion using an image generation AI. For example, the image generation AI can be trained using bone images and soft tissue images obtained by imaging and ES processing or forward projection of CT three-dimensional data, and new bone images and soft tissue images can be generated using the image generation AI. In this case, virtual monochromatic X-ray images and total thickness images can be generated from the new bone images and soft tissue images obtained by the image generation AI, and the pair of virtual monochromatic X-ray image and total thickness image can be used as training data. Note that when bone images and soft tissue images are used as output data of the training data, the new bone images and soft tissue images obtained by the image generation AI can be used as output data of the training data.

[0086] Furthermore, as a preprocessing of the training data, the processing unit 133 may normalize pixel values ​​of the output data of the training data between 0 and 1. By using such training data to train the machine learning model, it is possible to expect improvement in the accuracy of the training and stability of the training. Note that any known method may be used as a method for normalizing pixel values.

[0087] Next, in step S903, a learning process for the machine learning model is performed. Fig. 10 shows a block diagram of machine learning according to this embodiment. The ES image generation model of the present disclosure is a trained model generated by training (learning) based on a machine learning algorithm.

[0088] Here, we will briefly explain a general trained model. A trained model is a machine learning model that has been trained (learned) in advance using appropriate training data for an arbitrary machine learning algorithm. The training data consists of one or more pairs of input data and output data (ground truth data). The format and combination of input data and output data in the pairs that make up the training data may be suitable for a desired configuration, such as one being an image and the other being a numerical value, one being composed of a group of multiple images and the other being a character string, or both being images. Furthermore, a machine learning algorithm includes a method or system for performing training.

[0089] As shown in Fig. 10, in this embodiment, training data is loaded into a computer, and features such as regularities and patterns inherent in the training data are trained based on a machine learning algorithm to generate a trained model. Here, the machine learning algorithm includes deep learning techniques such as convolutional neural networks (CNNs). In deep learning techniques, different parameter settings for layers and nodes constituting a neural network may affect the degree to which trends trained using the training data can be reproduced in inference data.

[0090] Specifically, parameters in a CNN can include, for example, the filter kernel size, number of filters, stride value, and dilation value set for the convolution layer, as well as the number of nodes output by the fully connected layer. The parameter set and the number of training epochs can be set to values ​​suitable for the usage of the trained model based on the training data. For example, the parameter set and the number of epochs can be set based on the training data to enable the output of ES images such as total thickness images with higher accuracy.

[0091] Here is an example of one method for determining such parameter sets and the number of epochs. First, 70% of the pairs constituting the learning data are used for training, and the remaining 30% are randomly set as those for evaluation. Next, the training pairs are used to train the machine learning model, and at the end of each training epoch, the evaluation pairs are used to calculate a training evaluation value. The training evaluation value is, for example, the average value of a group of values ​​obtained by evaluating the output when input data constituting each pair is input to the machine learning model being trained, and the output data corresponding to the input data, using a loss function. Finally, the parameter set and the number of epochs that result in the smallest training evaluation value are determined as the parameter set and the number of epochs for the machine learning model.

[0092] In this way, by dividing the pairs that make up the learning data into those for training and those for evaluation and determining the number of epochs, it is possible to prevent the machine learning model from overfitting the training pairs.

[0093] An example of the configuration of a CNN related to an ES image generation model according to this embodiment will be described below with reference to Fig. 11. Fig. 11 shows an example of the configuration of an ES image model. The configuration shown in Fig. 11 is composed of a plurality of layer groups that are responsible for processing and outputting a group of input values. As shown in Fig. 11, the types of layers included in this configuration include a convolution layer, a downsampling layer, an upsampling layer, and a merging layer.

[0094] The convolution layer is a layer that performs convolution processing on a group of input values ​​according to parameters such as the set filter kernel size, the number of filters, the stride value, the dilation value, etc. Note that the number of dimensions of the filter kernel size may also be changed depending on the number of dimensions of the input image.

[0095] The downsampling layer is a layer that performs processing to reduce the number of output value groups to be less than the number of input value groups by thinning out or combining input value groups. Specifically, such processing includes, for example, max pooling processing.

[0096] An upsampling layer is a layer that performs processing to make the number of output values ​​greater than the number of input values ​​by duplicating input values ​​or adding values ​​interpolated from the input values, such as linear interpolation.

[0097] A synthesis layer is a layer that inputs a group of values, such as a group of output values ​​from a layer or a group of pixel values ​​that make up an image, from multiple sources and performs processing to combine them by concatenating or adding them.

[0098] In this configuration, the pixel values ​​constituting the input image Im1101 are output through a convolution processing block and then combined in a composition layer with the pixel values ​​constituting the input image Im1101. The combined pixel values ​​are then shaped into an ES image Im1102 in the final convolution layer.

[0099] It should be noted that different parameter settings for the layers and nodes that make up the neural network may affect the degree to which trends trained from learning data can be reproduced during inference. In other words, in many cases, appropriate parameters differ depending on the implementation form, so they can be changed to preferred values ​​as needed.

[0100] In addition to changing the parameters as described above, changing the configuration of the CNN can sometimes result in better CNN characteristics, such as outputting more accurate ES images, shortening processing time, or shortening the time required to train a machine learning model.

[0101] The CNN used in this embodiment is a U-net type machine learning model having an encoder function consisting of multiple layers including multiple downsampling layers and a decoder function consisting of multiple layers including multiple upsampling layers. That is, the CNN has a U-shaped structure having an encoder function and a decoder function. The U-net type machine learning model is configured (for example, using skip connections) so that position information (spatial information) obscured in multiple layers configured as encoders can be used in layers of the same dimension (layers corresponding to each other) in multiple layers configured as decoders.

[0102] Although not shown in the figure, as an example of a modification to the CNN configuration, for example, a batch normalization layer or an activation layer using a rectifier linear unit may be incorporated after the convolution layer.

[0103] The ES image generation model of this embodiment is configured to be stored in the storage unit 135 and executed by the processing unit 133, but these may also be provided in an external device (not shown) connected to the control device 103.

[0104] Here, a GPU can perform efficient calculations by processing a larger amount of data in parallel. Therefore, when performing learning multiple times using a machine learning algorithm such as deep learning, it is effective to use a GPU for processing. Therefore, in this embodiment, a GPU is used in addition to a CPU for processing by the processing unit 133, which functions as an example of a learning unit. Specifically, when executing a learning program including a learning model, the CPU and GPU work together to perform calculations to perform learning. Note that calculations may be performed only by the CPU or the GPU in the processing of the learning unit. Furthermore, the ES processing according to this embodiment may also be realized using a GPU, as with the learning unit. Note that if the trained model is provided in an external device, the processing unit 133 does not need to function as a learning unit.

[0105] The learning unit may also include an error detection unit and an update unit (not shown). The error detection unit obtains the error between correct data and output data output from the output layer of the neural network in response to input data input to the input layer. The error detection unit may use a loss function to calculate the error between the output data from the neural network and correct data. The update unit updates the connection weighting coefficients between the nodes of the neural network based on the error obtained by the error detection unit so as to reduce the error. This update unit updates the connection weighting coefficients using, for example, an error backpropagation method. The error backpropagation method is a technique for adjusting the connection weighting coefficients between the nodes of each neural network so as to reduce the error.

[0106] FIG. 12 shows the configuration of training data according to this embodiment. In this embodiment, both input and output data of the training data are image data. As described above, the input data of the training data are radiological images such as high-energy images, low-energy images, or virtual monochromatic X-ray images, and the output data are ES images such as total thickness images, bone images, or soft tissue images. As shown in FIG. 12, the training data is collective data including a plurality of these image pairs (training data 1 to n). Note that the training data according to this embodiment may include images generated by the data augmentation performed in step S902.

[0107] By using the trained model trained in this way, when a radiographic image is input to the trained model, the processing unit 133 can acquire an ES image such as a total thickness image corresponding to the input radiographic image. Furthermore, by constructing a trained model using such training data, many nonlinear arithmetic processes included in the ES processing can be included in inference using a machine learning algorithm such as deep learning.

[0108] (Image generation processing using a trained model) Next, an image generation process using the trained model described above will be described with reference to Fig. 13. Fig. 13 is a flowchart showing an example of an image generation process using the trained model. For ease of explanation, an example will be described below in which the input data of the training data and the target data are high-energy images, and the output data of the training data and the inference data of the trained model are total thickness images.

[0109] First, in step S1301, the acquisition unit 131 acquires a high-energy image acquired by normal imaging using the radiation imaging device 104. Note that the acquisition unit 131 may acquire the high-energy image from an external device connected to the control device 103. Furthermore, the acquired radiation image may be any radiation image that corresponds to the input data of the learning data, and may be, for example, a low-energy image.

[0110] Next, in step S1302, the processing unit 133 inputs the acquired high-energy image into the trained model and acquires a total thickness image output from the trained model. Here, FIG. 14 is a block diagram of inference processing using the trained model. When target data (input data) is input to the trained model, inference data according to the design of the trained model is output. At this time, the trained model outputs inference data that is likely to correspond to the target data, for example, according to a trend trained using the training data. In the example described here, the trained model outputs a total thickness image that is likely to correspond to the input high-energy image, according to the trained trend.

[0111] The combination of target data and inference data corresponds to the combination of input data and output data of the training data. Therefore, the target data may be a radiological image such as a high-energy image, a low-energy image, or an energy image obtained by imaging using the voltage of a virtual monochromatic X-ray image as the tube voltage during imaging, according to the input data of the training data. Furthermore, the inference data may be an ES image such as a total thickness image, a bone image, or a soft tissue image, according to the output data of the training data.

[0112] Here, with reference to FIGS. 15(a) and 15(b), the relationship between the input data and output data of the training data and the target data and inference data during inference will be described. FIG. 15(a) is a block diagram of a training process using exemplary training data, and FIG. 15(b) is a block diagram of an inference process using exemplary target data and inference data. As shown in FIG. 15(a), a trained model trained using high-energy images and total thickness images as training data trains features such as regularity and regularity inherent in the high-energy images and total thickness images. Therefore, as shown in FIG. 15(b), when a high-energy image is input as target data, the trained model can output a total thickness image as inference data that is likely to correspond to the target data according to the training trend.

[0113] As explained with reference to Figure 7, bones in the human body are surrounded by soft tissue, and when a soft-tissue image and a bone image are added together, the reduced bone thickness in the soft-tissue image is refilled. Therefore, the total thickness image obtained by adding the soft-tissue image and the bone image is an image in which the bone contrast has been removed, in other words, an image with fewer depicted object structures. Machine learning models are trained on the features of the training data, but when training is performed using data containing complex features, unnecessary artifacts due to the complex features may occur in the inference data. In contrast, training a machine learning model using images with fewer object structures can reduce artifacts due to the object structure that may occur in the inference data. Therefore, when a high-energy image captured as target data is input to a trained model trained using a high-energy image and a total thickness image as output data, a total thickness image with reduced artifacts can be output as inference data.

[0114] According to the processing of step S1302, an ES image can be generated from a single radiation image, such as a high-energy image or a low-energy image, using an ES image generation model based on machine learning. Therefore, when the control device 103 that performs such processing acquires an ES image using a trained model, there is no need to capture two images, a high-energy image and a low-energy image.

[0115] In contrast, in general ES processing, bone images and soft tissue images are generated using high-energy images and low-energy images. In order to acquire high-energy images and low-energy images with little time lag, such processing requires a detector equipped with a sample-and-hold circuit as shown in Fig. 2 and a radiation source capable of switching tube voltage as shown in Fig. 4. This requires large-scale hardware modifications to the radiation imaging device 104 and the radiation generator 101.

[0116] In contrast, the processing in step S1302 can be performed using a single radiographic image, which eliminates the need for the hardware modifications described above to the radiographic imaging system used for inference, thereby reducing the cost of the radiographic imaging system. Furthermore, the processing in step S1302 can be performed using a single radiographic image, which can also reduce the radiation exposure of the subject associated with capturing the radiographic image. Furthermore, since processing can be performed using a single radiographic image, there is an advantage in that motion artifacts due to misalignment that occur when processing is performed using two radiographic images are eliminated.

[0117] Next, in step S1303, the processing unit 133 performs post-processing on the ES image acquired in step S1302. Specifically, the processing unit 133 uses an approximate equation for ES processing to acquire an ES image of a type different from the total thickness image based on the high-energy image used as input data and the total thickness image acquired using the trained model. Here, the acquired ES image of a different type may be a material decomposition image such as a bone image or a soft tissue image, a virtual monochromatic X-ray image, or a radiological image of any energy. FIG. 16 is a block diagram showing an example of post-processing using the approximate equation for ES processing. In the example shown in FIG. 16, the processing unit 133 performs logarithmic transformation on the radiological image used as input data. Furthermore, the processing unit 133 generates an ES image of a type different from the total thickness image by performing weighted addition using the approximate equation for ES processing using the logarithm of the radiological image and the total thickness image acquired using the trained model.

[0118] In the explanation so far, the ES image generation model has learned the relationship between a radiographic image and a total thickness image, a radiographic image and a bone image, or a radiographic image and a soft tissue image. In such a configuration, an ES image generation model must be prepared for each ES image to be generated. However, the relationship shown in equation (4) holds for the pixel value X of the radiographic image, the bone thickness B, and the soft tissue thickness S. Here, for simplicity, let us consider the radiation spectrum N(E) when a radiographic image is acquired at any energy as the average energy E X Approximating this to a virtual monochromatic X-ray with a peak at , it is expressed by the following equation (11).

number

[0119] From equations (4) and (11), the pixel value X of a radiographic image acquired at any energy is calculated by the average energy E X The linear attenuation coefficient μ of bone in XB and the average energy E X The linear attenuation coefficient μ of soft tissue in XS Using this, it is expressed by the following equation (12).

number

[0120] Furthermore, the total thickness T is the sum of the bone thickness B and the soft tissue thickness S. Therefore, the bone thickness B and the soft tissue thickness S are expressed using the total thickness T by the following equation (13).

number

[0121] For simplicity, the spectrum N(E) of radiation when a radiological image is acquired at any energy is expressed as the average energy E Y In this case, the pixel value Y of a radiation image of any energy can be approximated by E Y The linear attenuation coefficient μ of bone in YB And, E Y The linear attenuation coefficient μ of soft tissue in YS Using this, it is expressed by the following equation (14).

number

[0122] Equation (13) multiplies the logarithm of the radiographic image and the total thickness image by a coefficient determined by the linear attenuation coefficient, and then adds them together. The same applies to equation (14). Therefore, as shown in Figure 16, ES images such as bone images, soft tissue images, virtual monochromatic X-ray images, and radiographic images of any energy can be obtained by weighted addition of the logarithm of the radiographic image and the total thickness image.

[0123] Furthermore, in this embodiment, a configuration was assumed in which the logarithms of the total thickness image generated by the ES image generation model and the radiographic image used as input data were weighted and added. However, this embodiment is not limited to such a configuration. As shown in FIG. 17 , a configuration may be adopted in which the logarithms of the bone image and the radiographic image generated by the ES image generation model are weighted and added to obtain ES images such as soft tissue images and total thickness images. In this case, the pixel values ​​S of the soft tissue image and the pixel values ​​T of the total thickness image are calculated using the logarithms of the bone image and the radiographic image as shown in the following equation (15).

number

[0124] Similarly, as shown in Fig. 18, a configuration may be adopted in which ES images such as bone images and total thickness images are obtained by weighted addition of the logarithms of the soft tissue image and radiological image generated by the ES image generation model. In this case, the method for calculating the pixel value B of the bone image and the pixel value T of the total thickness image using the logarithms of the soft tissue image and radiological image is as shown in the following equation (16).

number

[0125] Each of equations (13) to (16) is a weighted addition of the logarithm of a radiographic image and an ES image generated by the ES image generation model. Furthermore, the ES image obtained by this weighted addition is a different type of ES image from the ES image generated by the ES image generation model. Therefore, in the post-processing in step S1303, a different type of ES image from the ES image generated by the ES image generation model can be generated by weighting the logarithm of the radiographic image and the ES image generated by the ES image generation model.

[0126] In the above process, for simplicity, the radiation spectrum N(E) when a radiological image is acquired at any energy is expressed as the average energy E XThe ES image was approximated using virtual monochromatic X-rays with a peak at 1000 MHz. However, the inventors of the present application have found that calculations using approximation using virtual monochromatic X-rays may not result in accurate ES images. Therefore, a method for generating an ES image of a different type from the ES image generated by the ES image generation model without approximation using virtual monochromatic X-rays will be described.

[0127] The pixel value X of a radiographic image is calculated by the linear attenuation coefficient μ of bone at energy E. B (E) and the linear attenuation coefficient μ of soft tissue at energy E S Using (E), this is expressed as the following equation (17).

number

[0128] By solving the nonlinear equation of Equation (17), the thickness S of the soft tissue can be calculated from the bone thickness B and the pixel value X of the radiographic image. The method for solving Equation (17) may be the Newton-Raphson method, or an iterative method such as the least squares method or bisection method. When the Newton-Raphson method is used, for example, the value obtained by the approximation of Equation (17) can be used as the initial value. Alternatively, a table may be generated by calculating the soft tissue thickness S for various combinations of pixel value X of the radiographic image and bone thickness B, and the soft tissue thickness S may be calculated quickly by referencing the table.

[0129] Note that a similar process may be used to calculate the bone thickness S from the soft tissue thickness S and the pixel value X of the radiographic image. A total thickness image may also be obtained by adding a bone image or soft tissue image obtained using the ES image generation model and a soft tissue image or bone image obtained using Equation (17). A virtual monochromatic X-ray image may also be obtained using a bone image or soft tissue image obtained using the ES image generation model and a soft tissue image or bone image obtained using Equation (17). A bone image or soft tissue image may also be obtained using Equation (17) with the total thickness T and the pixel value X of the radiographic image. For example, by substituting (TS) for B or (TB) for S in Equation (17), a bone image or soft tissue image may be obtained using the total thickness image and the radiographic image. A bone image or soft tissue image may also be obtained using a table with the same process as above with the total thickness image and the radiographic image.

[0130] (A series of processes according to this embodiment) Next, a series of processes according to this embodiment, including image generation processing using the trained model, will be described with reference to Fig. 19 to Fig. 21. Fig. 19 is a flowchart showing an example of the series of processes according to this embodiment. Fig. 20 is a block diagram showing an example of the series of processes according to this embodiment.

[0131] In the image generation process using the trained model, it has been assumed that the pixel values ​​of the image output from the trained model are values ​​corresponding to physical units of length. However, for example, when the pixel values ​​of the output data of the training data are normalized between 0 and 1 or when an error occurs in the inference process using the trained model, the pixel values ​​of the image output from the trained model may not correspond to physical units of length.

[0132] In contrast, for example, in image generation processing using the trained model, calculations such as weighted addition are performed taking into consideration that the pixel values ​​of the image output from the trained model correspond to physical units of length. Similarly, when image analysis is performed using pixel values ​​of the image output from the trained model, calculations may be performed taking into consideration that the pixel values ​​of the image output from the trained model correspond to physical units of length. In these cases, if calculations are performed using pixel values ​​that do not correspond to physical units of length, appropriate calculation results may not be obtained, and the desired image or analysis results may not be obtained.

[0133] Therefore, the image processing device according to this embodiment performs unit conversion processing for converting pixel values ​​into physical lengths for an image output from a trained model, as shown in Fig. 20. Below, an example will be described in which processing is performed using a trained model that has been trained using images in which pixel values ​​do not correspond to physical units of length (e.g., cm) as output data from the training data, as an ES image generation model. For simplicity of explanation, an example will be described in which high-energy images are used as input data (target data) and bone images are used as output data from the ES image generation model.

[0134] First, in step S1901, the acquisition unit 131 acquires a radiographic image captured using the radiation imaging device 104. Note that the acquisition unit 131 may acquire the radiographic image from an external device connected to the control device 103.

[0135] In step S1902, the processing unit 133 inputs the radiographic image acquired in step S1901 into the ES image generation model, and acquires a bone image output from the ES image generation model. Here, the bone image output from the ES image generation model is an image whose pixel values ​​do not correspond to physical units of length, according to the training tendency.

[0136] Then, in step S1903, the processing unit 133 performs unit conversion processing on the bone image acquired in step S1902 to obtain a bone image whose pixel values ​​conform to cm, which is a physical unit of length. Specifically, the processing unit 133 multiplies the bone image generated in step S1902 by a scale value α, which is a unit conversion coefficient, to obtain a bone image whose pixel values ​​conform to cm, which is a physical unit of length. The method for determining the scale value α will be described later.

[0137] In step S1904, the processing unit 133 performs filtering such as low-pass filtering on the bone image after unit conversion acquired in step S1903. Note that by applying a low-pass filter (LPF), it is possible to reduce artifacts of high-frequency components generated by the ES image generation model.

[0138] In step S1905, the processing unit 133 performs logarithmic transformation processing on the radiation image acquired in step S1901. The logarithmic transformation processing may be performed by any known method.

[0139] In step S1906, the processing unit 133 acquires an ES image using an approximation equation for ES processing, based on the bone image after filtering obtained in step S1904 and the logarithm of the radiographic image obtained in step S1905. More specifically, the processing unit 133 performs weighted addition according to equation (15) using the bone image after filtering obtained in step S1904 and the logarithm of the radiographic image obtained in step S1905. This allows the processing unit 133 to acquire ES images such as soft tissue images and total thickness images in which pixel values ​​are in units of cm, which is a physical unit of length, as shown in FIG.

[0140] In step S1907, the display control unit 134 causes the display unit 120 to display the ES image acquired in step S1906. Note that the display control unit 134 may also cause the display unit 120 to display the analysis results of any analysis process performed by the processing unit 133 on the ES image acquired in step S1906. In addition, the control device 103 may transmit the ES image acquired in step S1906 and the analysis results of the ES image to an external device.

[0141] This process allows the pixel values ​​of an image output from a trained model to be converted into physical lengths, allowing the control device 103 to generate a desired image or perform desired image processing using an image whose pixel values ​​correspond to units of physical length.

[0142] In this embodiment, a high-energy image is used as input data, and a bone image is output from the ES image generation model. However, as described above, the input data and the inferred data output from the ES image generation model are not limited to this. The input data and the inferred data may be data corresponding to the training data. For example, the input data may be a low-energy image, or a radiological image such as an energy image obtained by imaging using the voltage of the virtual monochromatic X-ray image used as input data for the training data as the tube voltage during imaging. Furthermore, the ES image output from the ES image generation model and used in subsequent processing may be an ES image such as a soft-tissue image or a total thickness image.

[0143] When the total thickness image is used as the inference data, the strength of the low-pass filter processing can be increased, and the artifacts of the high-frequency components generated by the ES image generation model can be further reduced. Here, this effect will be described with reference to FIG.

[0144] Figure 21 is a diagram for explaining a total thickness image, and shows a schematic cross section of a human body. The surface of the human bone, i.e., the cortical bone, is covered with high-density bone. On the other hand, the inside of the bone, i.e., the trabecular bone, is formed of spongy bone. Therefore, radiographic images and bone images contain high-frequency components resulting from the structure of the cortical bone and trabecular bone.

[0145] However, the spaces between the trabeculae are filled with bone marrow, which is classified as soft tissue. Therefore, in the total thickness image, which is the sum of the bone image and soft tissue, the structures of the bone trabeculae and bone marrow cancel each other out, reducing the high-frequency components of the image. Therefore, in the total thickness image, signal components are less likely to be lost even if the strength of the low-pass filter applied is increased. Here, increasing the strength of the low-pass filter processing can further reduce high-frequency artifacts caused by the ES image generation model. Therefore, when using the total thickness image as inference data, the strength of the low-pass filter processing can be increased, further reducing high-frequency artifacts caused by the ES image generation model.

[0146] In this embodiment, the logarithmic conversion process in step S1905 is performed after the process in step S1904. However, the logarithmic conversion process in step S1905 may be performed between steps S1901 and S1906. The logarithmic conversion process in step S1905 may also be performed in parallel with the processes in steps S1901 to S1904.

[0147] Furthermore, in this embodiment, in step S1906, an ES image is acquired by weighted addition of the logarithm of the radiographic image and the unit-converted bone image according to equation (15) as an approximation formula for ES processing. On the other hand, if the image output from the ES image generation model is a soft tissue image, an ES image can be acquired by weighted addition of the logarithm of the radiographic image and the unit-converted soft tissue image according to equation (16) as an approximation formula for ES processing. Furthermore, if the image output from the ES image generation model is a bone image or a soft tissue image, the processing unit 133 may acquire an ES image by calculation using the radiographic image and an image obtained by converting the image output from the ES image generation model according to equation (17) as an approximation formula for ES processing. In this case, step S1905 may be omitted. Furthermore, if the image output from the ES image generation model is a total thickness image, the processing unit 133 can acquire an ES image by weighted addition of the logarithm of the radiographic image and the unit-converted total thickness image according to equation (13) or equation (14) as an approximation formula for ES processing. Furthermore, the processing unit 133 may acquire an ES image by performing a calculation using a nonlinear equation between the radiation image and the unit-converted total thickness image according to equation (17) as an approximation equation for ES processing. Note that the ES image acquired in step S1906 may be a material decomposition image such as a bone image or soft tissue image, a total thickness image, a virtual monochromatic X-ray image, or a radiation image of any energy.

[0148] (Method of determining unit conversion coefficients) Next, a method for determining a scale value α, which is a unit conversion coefficient used in the unit conversion process according to this embodiment, will be described with reference to Fig. 22 and Fig. 23. Fig. 22 is a flowchart showing an example of the process for determining the scale value α according to this embodiment. Fig. 23 is a block diagram showing an example of the process for determining the scale value α according to this embodiment. For ease of explanation, the following describes a case where a trained model obtained using bone images is used as output data of the training data.

[0149] As described above, in the total thickness image, the reduction in bone thickness in the soft tissue image is backfilled, thereby removing bone contrast. Therefore, for example, in a total thickness image of a subject consisting only of soft tissue and bone, such as a lower limb, the soft tissue (flesh) and bone are backfilled, resulting in an image containing continuous pixel values ​​(approximately uniform pixel values). In this regard, if the scale value α, which is a unit conversion coefficient, is appropriate, the total thickness image acquired using a bone image obtained by converting pixel values ​​into physical lengths using the scale value α and a radiographic image will contain continuous pixel values. Therefore, the method for determining the unit conversion coefficient according to this embodiment determines the scale value α by utilizing the continuity of the total thickness image.

[0150] First, in step S2201, the acquisition unit 131 acquires a radiographic image captured using the radiation imaging device 104. Note that the acquisition unit 131 may acquire the radiographic image from an external device connected to the control device 103.

[0151] In step S2202, the processing unit 133 inputs the radiographic image acquired in step S2201 into the ES image generation model, and acquires a bone image output from the ES image generation model. Here, the bone image output from the ES image generation model is an image whose pixel values ​​do not correspond to physical units of length, according to the training tendency.

[0152] In step S2203, the processing unit 133 applies a plurality of scale values ​​α, which are arbitrarily set in advance, to the bone image acquired in step S2202. i 23, the processing unit 133 multiplies each of the scale values ​​α (α1, α2, . . . ) (candidates for the scale values) by α. i A plurality of bone images Im B [α i ] to get.

[0153] In step S2204, the processing unit 133 combines the radiographic image acquired in step S2201 and the multiple bone images Im acquired in step S2203. B [α i] and solve the nonlinear equation of Equation (17) for each pixel. i A plurality of soft tissue images Im corresponding to each of the S [α i ]. Then, the processing unit 133 obtains the scale value α i The bone images Im B [α i ] and soft tissue imaging Im S [α i ] are added together and the scale value α i The total thickness image Im T [α i For example, the processing unit 133 acquires the bone image Im corresponding to the scale value α1. B [α1] and soft tissue image Im S Add [α1] to the total thickness image Im T [α1]. The processing unit 133 also acquires the bone image Im corresponding to the scale value α2. B [α2] and soft tissue image Im S Add [α2] to the total thickness image Im T [α2]. Through this process, the processing unit 133 obtains the scale value α i A plurality of total thickness images Im T [α i ] to get.

[0154] By devising a method for solving the nonlinear equation of Equation (17), the processing unit 133 can obtain the radiographic image and the bone image Im corresponding to the scale value α1. B [α i ] and the scale value α i The total thickness image Im T [α i For example, since T=B+S, a solution to the nonlinear equation of equation (17) can be devised by substituting (TB) for S in equation (17).

[0155] In step S2205, the processing unit 133 converts the plurality of total thickness images Im acquired in step S2204 into T [α i] are subjected to a high-pass filter, and multiple filtered total thickness images Im T [α i ]' is obtained. Note that a band-pass filter may be applied instead of a high-pass filter. Here, the band-pass filter may be set to pass a frequency band so that the band-pass filter is applied to the total thickness image, making it easier to evaluate the continuity of the total thickness image in subsequent processing. For example, the frequency band to pass may be set to a frequency band corresponding to high-frequency components that are reduced in an appropriate total thickness image.

[0156] In step S2206, the processing unit 133 generates a plurality of filtered total thickness images Im T [α i For example, the processing unit 133 performs statistical processing on each of the total thickness images Im T [α i ]' are averaged to obtain multiple total thickness images Im T [α i ]' corresponding to multiple average values ​​A[α i ] is acquired. Note that the statistical processing is not limited to averaging processing, and the statistical value is not limited to the average value. The acquired statistical value may be, for example, a median or a maximum value, and the statistical processing may be any processing appropriate for the acquired statistical value.

[0157] In step S2207, the processing unit 133 determines the scale value corresponding to the smallest statistical value among the plurality of statistical values ​​obtained in step S2206 as the scale value α. For example, the processing unit 133 determines the plurality of average values ​​A[α i ], the scale value corresponding to the smallest average value is determined as the scale value α.

[0158] According to this process, the scale value α, which is a unit conversion coefficient, can be determined by utilizing the continuity of the total thickness image. In addition, by calculating statistics for the high-frequency components of the total thickness image and determining the scale value corresponding to the minimum statistical value, it is possible to search for a total thickness image in which the high-frequency components are appropriately reduced, and to search for an appropriate scale value corresponding to the total thickness image.

[0159] In step S2206, the processing unit 133 may set an arbitrary region of interest (ROI) for each total thickness image and calculate statistical values ​​in the ROI.

[0160] In this embodiment, the scale value α is searched for using all of the multiple preset scale values. However, the method for searching for the scale value α is not limited to this. For example, the scale value α may be searched for in stages by performing a search similar to the above using scale values ​​that are far apart, and then repeating the same search using scale values ​​close to the searched scale value. Furthermore, the scale value α may be searched for by increasing or decreasing the scale value from a preset initial value. In this case, for example, the processes of steps S2203 to S2206 may be repeated while increasing or decreasing the scale value from the initial value, and the scale value when the statistical value acquired in step S2206 becomes smaller than the threshold may be determined as the scale value α.

[0161] In this embodiment, for simplicity of explanation, a case has been described in which the image output from the ES image generation model used in the process of determining the scale value α is a bone image. However, the ES image generation model used in the process of determining the scale value α is not limited to an ES image generation model that outputs a bone image. For example, the ES image generation model used in the process of determining the scale value α may be a trained model obtained using a soft-tissue image as output data of the training data, and a soft-tissue image may be output from the ES image generation model. In this case, the processing unit 133 can generate a total thickness image using the soft-tissue image by processes similar to those in steps S2203 to S2207, and then determine the scale value α based on the total thickness image.

[0162] Furthermore, the ES image generation model used in the process of determining the scale value α may be a trained model obtained using a total thickness image as output data of the training data, or the total thickness image may be output from the ES image generation model. In this case, the processing unit 133 acquires a bone image or a soft tissue image by weighted addition according to Equation (13) or Equation (14) using, for example, the total thickness image output from the ES image generation model and the logarithm of the radiographic image acquired in step S2201. The processing unit 133 may also acquire a bone image or a soft tissue image by calculation using a nonlinear equation according to Equation (17) using the total thickness image output from the ES image generation model and the radiographic image acquired in step S2201. In this case, for example, in Equation (17), the nonlinear equation may be solved after substituting (TB) for S or (TS) for B. Thereafter, the processing unit 133 can generate a total thickness image using the acquired bone image or soft tissue image by processing similar to steps S2203 to S2207, and then determine the scale value α based on the generated total thickness image.

[0163] In step S2204, the total thickness image is acquired using equation (17), but the total thickness image may also be acquired using equation (15). Also, when a soft tissue image is output from the ES image generation model, the total thickness image may be acquired using equation (16) in step S2204. When the total thickness image is acquired using the nonlinear equation of equation (17) in step S2204, approximation using virtual monochromatic X-rays is not used, and therefore a more accurate total thickness image can be acquired.

[0164] As described above, the radiation imaging system according to this embodiment includes the radiation imaging device 104 that detects radiation irradiated from the radiation generation device 101, and the control device 103 that is communicably connected to the radiation imaging device 104. The control device 103 can function as an example of an image processing device that performs image processing.

[0165] The control device 103 includes an acquisition unit 131 and a processing unit 133. The acquisition unit 131 functions as an example of an acquisition unit that acquires a radiographic image. The processing unit 133 functions as an example of a processing unit that acquires an energy subtraction image by inputting the acquired radiographic image into a trained model. The processing unit 133 also applies unit conversion processing to the acquired energy subtraction image, converting pixel values ​​into physical lengths.

[0166] With this configuration, the control device 103 according to this embodiment can convert pixel values ​​of an image output from an ES image generation model into physical lengths. This prevents a situation in which appropriate calculation results cannot be obtained using the image output from the ES image generation model due to pixel values ​​not conforming to physical length units. Therefore, the control device 103 can obtain desired images and analysis results using the image output from the ES image generation model.

[0167] The unit conversion process according to this embodiment may include multiplying the energy subtraction image acquired using the trained model by a unit conversion coefficient, which may be acquired by the following process.

[0168] First, a radiographic image is input to the trained model, and a total thickness image showing the sum of the thicknesses of different materials in the subject is generated using an image obtained by performing unit conversion processing on the energy subtraction image output from the trained model using candidate unit conversion coefficients. Then, the unit conversion coefficient can be obtained based on the candidate unit conversion coefficients according to the statistical values ​​of the feature quantities of the total thickness image. The unit conversion coefficient can be obtained based on the candidate unit conversion coefficient corresponding to the total thickness image with the lowest statistical value among multiple total thickness images generated using multiple candidate unit conversion coefficients. This process allows the unit conversion coefficient to be obtained with high accuracy by utilizing the continuity of the total thickness image.

[0169] The total thickness image can be generated by calculation using a nonlinear equation that uses a radiographic image input to the trained model and an image that has undergone unit conversion processing using candidate unit conversion coefficients. Alternatively, the total thickness image may be generated by weighted addition that uses the logarithm of the radiographic image input to the trained model and an image that has undergone unit conversion processing using candidate unit conversion coefficients.

[0170] The feature of the total thickness image may be obtained by applying a high-pass filter to the total thickness image. In this case, the high-frequency components reduced in the total thickness image can be used as the feature, which makes it easier to evaluate the continuity of the total thickness image.

[0171] The statistical value of the feature amount of the total thickness image may be any one of the average, median, and maximum values ​​of the feature amount of the total thickness image, or may be a statistical value of the feature amount of a predetermined region (ROI) in the total thickness image.

[0172] The energy subtraction image acquired using the trained model may be a material decomposition image showing the thickness of a material in the subject or a total thickness image showing the sum of the thicknesses of different materials in the subject. Here, the material decomposition image may be, for example, a bone image or a soft tissue image.

[0173] Furthermore, the processing unit 133 can use the energy subtraction image to which the unit conversion process has been applied and the acquired radiographic image to acquire a different type of energy subtraction image from the energy subtraction image. More specifically, the processing unit 133 can acquire a different type of energy subtraction image from the energy subtraction image by weighted addition using the energy subtraction image to which the unit conversion process has been applied and the logarithm of the acquired radiographic image. The processing unit 133 may also acquire a different type of energy subtraction image from the energy subtraction image to which the unit conversion process has been applied and the acquired radiographic image by calculation using a nonlinear equation.

[0174] Furthermore, with the above configuration, a desired ES image can be generated using a single radiographic image. This eliminates the need for significant hardware modifications to the radiographic imaging system used for inference using a trained model, thereby reducing the cost of the radiographic imaging system. Furthermore, since processing can be performed using a single radiographic image, it is also possible to reduce the amount of radiation exposure of the subject associated with capturing the radiographic image.

[0175] Furthermore, the processing unit 133 can use an energy subtraction image to which a low-pass filter and unit conversion process have been applied to acquire a different type of energy subtraction image from the energy subtraction image. In this case, it is possible to reduce artifacts due to high-frequency components in the image output by the trained model.

[0176] The acquired different types of energy subtraction images may be any of a material decomposition image showing the thickness of a material in the subject, a total thickness image showing the sum of the thicknesses of different materials in the subject, a virtual monochromatic X-ray image, and a radiation image of any energy. Here, the material decomposition image may be, for example, a bone image or a soft tissue image.

[0177] (Variation 1) In the first embodiment, the unit conversion process involves multiplying the ES image obtained using the ES image generation model by a scale value α. However, the method of unit conversion is not limited to this. For example, the unit conversion process may involve converting the ES image obtained using the ES image generation model so that the mean and variance of the ES image obtained using the ES image generation model match the target mean and variance. Any known method may be used to convert the image.

[0178] The target mean and variance may be set in the same manner as the method for setting the scale value α described with reference to FIGS. 22 and 23 in the first embodiment. For example, the processing unit 133 acquires a radiographic image in step S2201, and generates an ES image from the radiographic image using an ES image generation model in step S2202. Thereafter, in step S2203, the processing unit 133 converts the ES image output from the ES image generation model so that the mean and variance of the ES image become each of a plurality of predetermined means and variances. The subsequent processing may be the same as in the first embodiment. However, in step S2207, instead of the scale value α, the mean and variance corresponding to the smallest statistical value among the statistical values ​​calculated in step S2206 are set as the target mean and variance.

[0179] The method of unit conversion may be determined depending on the ES image obtained using the trained model. For example, if the lower limit of pixel values ​​in the ES image obtained using the trained model is 0, unit conversion using the scale value α may be performed.

[0180] As described above, the unit conversion process according to this embodiment can include converting the acquired energy subtraction image so that the mean and variance of the acquired energy subtraction image become the mean and variance of the target, where the mean and variance of the target can be obtained by the following process:

[0181] A radiological image is input to the trained model, and an energy subtraction image output from the trained model is subjected to unit conversion processing using candidate target means and variances to generate a total thickness image indicating the sum of the thicknesses of different materials in the subject. The total thickness image can then be obtained based on the candidates according to the statistical values ​​of the features of the total thickness image. The target mean and variance can be obtained based on the candidate corresponding to the total thickness image with the lowest statistical value among multiple total thickness images generated using multiple candidate target means and variances. This processing allows the target mean and variance to be obtained with high accuracy by utilizing the continuity of the total thickness image.

[0182] The total thickness image can be generated by calculation using a nonlinear equation that uses a radiographic image input to the trained model and an image that has undergone unit conversion processing using candidates for the target mean and variance. Alternatively, the total thickness image may be generated by weighted addition that uses the logarithm of the radiographic image input to the trained model and an image that has undergone unit conversion processing using candidates for the target mean and variance.

[0183] (Variation 2) In the first embodiment, a configuration has been described in which radiographic images and bone images are prepared as training data to train the ES image generation model. In such a case, a total thickness image may be generated using the bone images output from the ES image generation model, and low-pass filtering or the like may be performed on the total thickness image.

[0184] 24 and 25, a second modification will be described in which, when bone images are used as output data of the learning data, a total thickness image is generated using the bone images output from the ES image generation model, and an ES image is generated based on the total thickness image. Fig. 24 is a flowchart showing an example of a series of processes according to this modification. Fig. 25 is a block diagram showing an example of a series of processes according to this modification.

[0185] The series of processes according to this modification differs from the series of processes according to the first embodiment in that steps S2401 and S2402 are added after step S1903, and step S1905 is deleted. The series of processes according to this embodiment will be described below, focusing on the differences from the first embodiment.

[0186] The processing of steps S1901 to S1903 will not be described because it is the same as the processing according to the first embodiment. In step S1903, when a unit-converted bone image is acquired, the processing proceeds to step S2401.

[0187] In step S2401, the processing unit 133 performs logarithmic conversion processing on the radiation image acquired in step S1901 by the same processing as in step S1905 according to the first embodiment.

[0188] In step S2402, the processing unit 133 performs weighted addition of the bone image after unit conversion obtained in step S1903 and the logarithm of the radiation image obtained in step S2401 to obtain a total thickness image, as shown in Fig. 25. Equation (15) can be used for the weighted addition.

[0189] Thereafter, in step S1904, the processing unit 133 applies a low-pass filter to the total thickness image obtained in step S2402. After the low-pass filter is applied to the total thickness image in step S1904, the process proceeds to step S1906. The process of step S1906 is the same as the process of step S1906 according to the first embodiment, except that the logarithm of the radiographic image obtained in step S2401 is used instead of the logarithm of the radiographic image obtained in step S1905. The subsequent processes may be the same as the series of processes according to the first embodiment.

[0190] In this modification, when generating a bone image using an ES image generation model, the ES image is generated after passing through a total thickness image. With this configuration, the low-pass filter is applied to the total thickness image, so that the strength of the low-pass filter processing can be increased as described above, and high-frequency artifacts caused by the ES image generation model can be further reduced.

[0191] It is also possible to prepare radiographic images and soft-tissue images to train the ES image generation model, generate a total thickness image using the soft-tissue images output from the ES image generation model, and then perform low-pass filtering on the total thickness image. Fig. 26 is a block diagram showing an example of a series of processes according to this modified example when soft-tissue images are used as output data of the training data.

[0192] In this case, the same processing as the series of processing shown in Fig. 24 may be performed. However, in step S2402, as shown in Fig. 26, the processing unit 133 performs weighted addition of the soft-tissue image after unit conversion obtained in step S1903 and the logarithm of the radiographic image obtained in step S2201 to obtain a total thickness image. Equation (16) can be used for the weighted addition. The subsequent processing may be the same as the series of processing shown in Fig. 24.

[0193] In this series of processes, when generating a soft-tissue image using an ES image generation model, the ES image is generated after passing through a total thickness image. Even in this case, the low-pass filter is applied to the total thickness image, so the strength of the low-pass filter processing can be increased as described above, and high-frequency component artifacts caused by the ES image generation model can be further reduced.

[0194] In this modification, the logarithmic conversion process in step S2401 is performed after the process of step S1903. However, the logarithmic conversion process in step S2401 may be performed between steps S1901 and S2402. Furthermore, the logarithmic conversion process in step S2401 may be performed in parallel with the processes from steps S1901 to S1903.

[0195] In this modification, in step S2402, the processing unit 133 acquires a total thickness image by weighted addition of the logarithm of the radiographic image and the image output from the ES image generation model according to equation (15) or equation (16) as an approximation equation for ES processing. Alternatively, the processing unit 133 may acquire a total thickness image using equation (17) as an approximation equation for ES processing. More specifically, the processing unit 133 may acquire a soft tissue image or a bone image by performing a calculation using the radiographic image and the bone image or soft tissue image output from the ES image generation model according to equation (17). The processing unit 133 can acquire a total thickness image by adding the bone image or soft tissue image output from the ES image generation model and the soft tissue image or bone image acquired according to the approximation equation for ES processing. Alternatively, the processing unit 133 may directly acquire a total thickness image by devising a solution using the Newton-Raphson method for equation (17). In these cases, step S2401 may be omitted.

[0196] As described above, the processing unit 133 according to this modification can acquire a total thickness image indicating the sum of the thicknesses of different materials in the subject, using an energy subtraction image to which unit conversion processing has been applied and an acquired radiographic image. The processing unit 133 can also apply a low-pass filter to the acquired total thickness image. Furthermore, the processing unit 133 can acquire a type of energy subtraction image different from the total thickness image by weighted addition using the total thickness image to which the low-pass filter has been applied and the logarithm of the acquired radiographic image. The processing unit 133 may also acquire a type of energy subtraction image different from the total thickness image by performing a nonlinear equation calculation using the total thickness image to which the low-pass filter has been applied and the acquired radiographic image. In this case, the different type of energy subtraction image may be any of a material decomposition image indicating the thickness of materials in the subject, a virtual monochromatic X-ray image, and a radiographic image of any energy.

[0197] This configuration can also achieve the same effects as in the first embodiment. In addition, since a low-pass filter is applied to the total thickness image, the filter strength can be improved, and artifacts in the image output by the trained model can be further reduced.

[0198] (Variation 3) Note that when images clipped by data augmentation are used as training data for a trained model, the processing unit 133 can clip (divide) the radiographic image to be input to the ES image generation model to make it an image size corresponding to the training data. The processing unit 133 can then connect the images together to make the ES image output from the ES image generation model by inputting the clipped radiographic image the same size as before clipping. Below, an example will be described in which the input data and target data of the training data are high-energy images, and the output data of the training data and the inference data of the trained model are total thickness images. Here, FIG. 27(a) is a block diagram showing an example of processing using clipped images during training, and FIG. 27(b) is a block diagram showing an example of processing using clipped images during inference.

[0199] When performing such processing, tile-like artifacts may occur due to differences in levels at the joints between images. Therefore, the processing unit 133 may provide overlapping portions (overlapping portions) between the images when clipping the radiographic images, and join the overlapping portions of the total thickness image output from the ES image generation model by weighting and adding them together. For example, the weighting may be performed based on the coordinates of the pixels in the overlapping portions of the clipped radiographic images.

[0200] As described above, the processing unit 133 according to this modification can divide an acquired radiographic image into multiple images including overlapping regions and input the images to the trained model. Furthermore, the processing unit 133 can acquire an energy subtraction image by joining multiple images output from the trained model based on the overlapping regions. This can reduce the occurrence of tile-like artifacts due to steps at the joints between images.

[0201] In the above embodiment and modified examples, the scale value α, which is a unit conversion coefficient, and the target mean and variance are set by the control device 103. However, the acquisition unit 131 of the control device 103 may acquire the scale value α, which is a unit conversion coefficient, and the target mean and variance from an external device (not shown) connected to the control device 103. In this case, the scale value α, which is a unit conversion coefficient, and the target mean and variance may be any that correspond to the ES image generation model used in the above series of processes.

[0202] Note that the post-processing of the processing using the trained model according to the above-described embodiment and modified examples is not limited to the above-described processing. For example, it may include tone conversion processing, any image quality improvement processing, any resolution improvement processing, any image conversion processing, etc. Furthermore, in the post-processing of the ES image generation processing according to the above-described embodiment and modified examples, filtering processing using a low-pass filter or the like is performed, but filtering may not be performed. Furthermore, the filter used in filtering processing may be a band-pass filter that passes signals in a desired frequency band.

[0203] In addition, in the trained models described in the above embodiments and variants, it is considered that the magnitude of the brightness values ​​of the input data image, the order and gradient of the bright and dark areas, position, distribution, continuity, etc. are extracted as part of the features and used in the inference process.

[0204] (Other Examples) The present disclosure can also be realized by providing a program that implements one or more functions of the above-described embodiments to a system or device via a network or a storage medium, and having one or more processors in the computer of the system or device read and execute the program. It can also be realized by a circuit (e.g., ASIC) that implements one or more functions. A computer may have one or more processors or circuits, and may include multiple separate computers or a network of multiple separate processors or circuits to read and execute computer-executable instructions.

[0205] The processor or circuitry may include a central processing unit (CPU), a microprocessing unit (MPU), a graphics processing unit (GPU), an application specific integrated circuit (ASIC), or a field programmable gateway (FPGA). The processor or circuitry may also include a digital signal processor (DSP), a data flow processor (DFP), or a neural processing unit (NPU).

[0206] The above disclosure includes the following configurations, methods, and programs. (Configuration 1) an acquisition unit for acquiring a radiation image; a processing unit that inputs the acquired radiographic image into a trained model to acquire an energy subtraction image; Equipped with The processing unit applies unit conversion processing to the acquired energy subtraction image to convert pixel values ​​into physical lengths. (Configuration 2) 2. The image processing device according to claim 1, wherein the unit conversion process includes a process of multiplying the acquired energy subtraction image by a unit conversion coefficient, or a process of converting the acquired energy subtraction image so that a mean and a variance of the acquired energy subtraction image become target mean and variance. (Configuration 3) The unit conversion factor or the target mean and variance are: A radiation image is input to the trained model, and an energy subtraction image output from the trained model is subjected to unit conversion processing using the unit conversion coefficient or the target mean and variance candidates, and a total thickness image indicating the sum of the thicknesses of different materials in the subject is generated using the image; 3. The image processing device according to configuration 2, wherein the candidate is acquired according to a statistical value of a feature amount of the total thickness image. (Configuration 4) The image processing device according to configuration 3, wherein the unit conversion coefficient or the target mean and variance are obtained based on a candidate corresponding to the total thickness image with the lowest statistical value among a plurality of total thickness images generated using a plurality of candidates for the unit conversion coefficient or the target mean and variance. (Configuration 5) 5. The image processing device according to configuration 3 or 4, wherein the total thickness image is generated by calculation using a nonlinear equation that uses a radiographic image input to the trained model and an image that has undergone unit conversion processing using the candidate. (Configuration 6) 5. The image processing device according to configuration 3 or 4, wherein the total thickness image is generated by weighted addition using the logarithm of the radiological image input to the trained model and an image subjected to unit conversion processing using the candidate. (Configuration 7) 7. The image processing device according to any one of configurations 3 to 6, wherein the feature amount is obtained by applying a high-pass filter to the total thickness image. (Configuration 8) 8. The image processing device according to any one of configurations 3 to 7, wherein the statistical value is one of an average value, a median value, and a maximum value of the feature amount of the total thickness image. (Configuration 9) 9. The image processing device according to any one of configurations 3 to 8, wherein the statistical value is a statistical value of a feature amount of a predetermined region in the total thickness image. (Configuration 10) The processing unit Dividing the acquired radiographic image into a plurality of images including overlapping regions and inputting the images into the trained model; 10. The image processing device according to any one of configurations 1 to 9, wherein a plurality of images output from the trained model are joined together based on the overlapping region to obtain an energy subtraction image. (Configuration 11) 11. The image processing device according to any one of configurations 1 to 10, wherein the energy subtraction image is a material decomposition image showing the thickness of a material of the subject or a total thickness image showing the sum of thicknesses of different materials of the subject. (Configuration 12) 12. The image processing device according to any one of configurations 1 to 11, wherein the processing unit uses the energy subtraction image to which the unit conversion processing has been applied and the acquired radiation image to acquire an energy subtraction image of a different type from the energy subtraction image. (Configuration 13) The processing unit weighted addition using the energy subtraction image to which the unit conversion processing has been applied and the logarithm of the acquired radiation image; or 13. The image processing device according to claim 12, wherein an energy subtraction image of a different type from the energy subtraction image is obtained by performing a calculation using a nonlinear equation that uses the energy subtraction image to which the unit conversion processing has been applied and the obtained radiation image. (Configuration 14) 14. The image processing device according to claim 12, wherein the processing unit uses an energy subtraction image to which a low-pass filter has been applied and to which the unit conversion process has been applied to obtain an energy subtraction image of a different type from the energy subtraction image. (Configuration 15) The image processing device according to any one of configurations 12 to 14, wherein the different types of energy subtraction images are any of a material decomposition image showing the thickness of a material in the subject, a total thickness image showing the sum of the thicknesses of different materials in the subject, a virtual monochromatic X-ray image, and a radiation image of any energy. (Configuration 16) The processing unit obtaining a total thickness image indicating the sum of thicknesses of different materials in the subject using the energy subtraction image to which the unit conversion process has been applied and the obtained radiation image; applying a low-pass filter to the acquired total thickness image; weighted addition using the low-pass filtered total thickness image and the logarithm of the acquired radiographic image; or 12. The image processing device according to any one of configurations 1 to 11, wherein an energy subtraction image of a type different from the total thickness image is acquired by performing a calculation using a nonlinear equation that uses the total thickness image to which the low-pass filter has been applied and the acquired radiation image. (Configuration 17) 17. The image processing device according to claim 16, wherein the different types of energy subtraction images are any of a material decomposition image showing the thickness of a material in the subject, a virtual monochromatic X-ray image, and a radiation image of any energy. (Configuration 18) a radiation imaging device for detecting radiation; the image processing device according to any one of configurations 1 to 17, which is communicably connected to the radiation imaging device; A radiation imaging system comprising: (Method 1) obtaining a radiological image; inputting the acquired radiographic image into a trained model to acquire an energy subtraction image; applying a unit conversion process to the acquired energy subtraction image to convert pixel values ​​into physical lengths; An image processing method comprising: (Program 1) A program that, when executed by a computer, causes the computer to perform the image processing method described in Method 1.

[0207] Although the present disclosure has been described above with reference to embodiments and modifications, the present disclosure is not limited to the above embodiments and modifications. Inventions modified within the scope of the present disclosure and inventions equivalent to the present disclosure are also included in the present disclosure. Furthermore, the above-described embodiments and modifications can be combined as appropriate within the scope of the present disclosure. [Explanation of symbols]

[0208] 103: Control device (image processing device) 131: Acquisition Department 133: Processing section

Claims

1. an acquisition unit for acquiring a radiation image; a processing unit that inputs the acquired radiographic image into a trained model to acquire an energy subtraction image; Equipped with The processing unit applies unit conversion processing to the acquired energy subtraction image to convert pixel values ​​into physical lengths.

2. 2. The image processing device according to claim 1, wherein the unit conversion process includes a process of multiplying the acquired energy subtraction image by a unit conversion coefficient, or a process of converting the acquired energy subtraction image so that the mean and variance of the acquired energy subtraction image become target mean and variance.

3. The unit conversion factor or the target mean and variance are: A radiation image is input to the trained model, and an energy subtraction image output from the trained model is subjected to unit conversion processing using the unit conversion coefficient or the target mean and variance candidates, and a total thickness image indicating the sum of the thicknesses of different materials in the subject is generated using the image; The image processing device according to claim 2 , wherein the candidate is acquired according to a statistical value of a feature amount of the total thickness image.

4. The image processing device of claim 3, wherein the mean and variance of the unit conversion coefficient or the target are obtained based on the candidate corresponding to the total thickness image with the lowest statistical value among multiple total thickness images generated using multiple candidates for the mean and variance of the unit conversion coefficient or the target.

5. The image processing device according to claim 3 , wherein the total thickness image is generated by calculation using a nonlinear equation that uses a radiographic image input to the trained model and an image that has undergone unit conversion processing using the candidate.

6. The image processing device according to claim 3 , wherein the total thickness image is generated by weighted addition using the logarithm of the radiographic image input to the trained model and an image subjected to unit conversion processing using the candidate.

7. The image processing device according to claim 3 , wherein the feature amount is obtained by applying a high-pass filter to the total thickness image.

8. The image processing device according to claim 3 , wherein the statistical value is one of an average value, a median value, and a maximum value of the feature amount of the total thickness image.

9. The image processing device according to claim 3 , wherein the statistical value is a statistical value of a feature amount of a predetermined region in the total thickness image.

10. The processing unit Dividing the acquired radiographic image into a plurality of images including overlapping regions and inputting the images into the trained model; The image processing device according to claim 1 , wherein the image processing device acquires an energy subtraction image by stitching together a plurality of images output from the trained model based on the overlapping region.

11. The image processing apparatus according to claim 1 , wherein the energy subtraction image is a material decomposition image showing the thickness of a material of the object or a total thickness image showing the sum of thicknesses of different materials of the object.

12. The image processing device according to claim 1 , wherein the processing unit acquires an energy subtraction image of a different type from the energy subtraction image to which the unit conversion processing has been applied, using the acquired radiation image.

13. The processing unit weighted addition using the energy subtraction image to which the unit conversion processing has been applied and the logarithm of the acquired radiation image; or 13. The image processing device according to claim 12, wherein an energy subtraction image of a type different from the energy subtraction image is obtained by performing a calculation using a nonlinear equation that uses the energy subtraction image to which the unit conversion processing has been applied and the obtained radiation image.

14. The image processing device according to claim 12 , wherein the processing unit uses an energy subtraction image to which a low-pass filter has been applied and to which the unit conversion processing has been applied to obtain an energy subtraction image of a different type from the energy subtraction image.

15. 13. The image processing device according to claim 12, wherein the different types of energy subtraction images are any of a material decomposition image showing the thickness of a material in the object, a total thickness image showing the sum of thicknesses of different materials in the object, a virtual monochromatic X-ray image, and a radiation image of any energy.

16. The processing unit obtaining a total thickness image indicating the sum of thicknesses of different materials in the subject using the energy subtraction image to which the unit conversion process has been applied and the obtained radiation image; applying a low-pass filter to the acquired total thickness image; weighted addition using the low-pass filtered total thickness image and the logarithm of the acquired radiographic image; or The image processing device according to claim 1 , further comprising: a calculation using a nonlinear equation that uses the total thickness image to which the low-pass filter has been applied and the acquired radiation image to acquire an energy subtraction image of a type different from the total thickness image.

17. The image processing apparatus according to claim 16, wherein the different types of energy subtraction images are any of a material decomposition image showing the thickness of a material in the subject, a virtual monochromatic X-ray image, and a radiation image of any energy.

18. a radiation imaging device for detecting radiation; an image processing device according to any one of claims 1 to 17, which is communicably connected to the radiation imaging device; A radiation imaging system comprising:

19. obtaining a radiological image; inputting the acquired radiographic image into a trained model to acquire an energy subtraction image; applying a unit conversion process to the acquired energy subtraction image to convert pixel values ​​into physical lengths; An image processing method comprising:

20. A program that, when executed by a computer, causes the computer to execute the image processing method according to claim 19.

Citation Information

Patent Citations

  • Radiation imaging system, imaging control device, and method

    JP2019162358A

  • Image processing device, method for processing image, and program

    JP2023120851A