Recycled metal intelligent classification method based on spark stream video dynamic analysis
Through the intelligent classification method based on spark stream video dynamic analysis, the problem of inefficiency of traditional metal classification methods in complex industrial scenarios is solved, efficient and real-time online metal classification is achieved, and recycling costs are reduced.
Patent Information
- Application Number
- CN202510670471.1
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-05-23
- Publication Date
- 2025-06-20
- Estimated Expiration
- 2045-05-23
AI Technical Summary
Traditional metal classification methods are difficult to meet real-time requirements in complex industrial scenarios, especially in high-frequency metal recycling assembly lines, resulting in low efficiency and high error rates.
Using an intelligent classification method based on dynamic analysis of spark stream videos, a high-speed camera is used to collect spark stream videos, and a YOLOv8 model is used to detect the spark area, extract the dynamic characteristics of the spark, and a metal classification model is constructed based on the bidirectional LSTM model to realize real-time online classification.
It significantly improves the classification efficiency of different materials in scrap metals, reduces recycling costs, adapts to complex industrial scenarios, and improves the robustness of spark area detection.
Smart Images

Figure CN120182736A_ABST
Abstract
Description
Technical Field
[0001] The present invention relates to the field of image recognition, and particularly to an intelligent classification method for recycling and reusing metals based on dynamic analysis of spark flow videos. Background Art
[0002] In the process of waste resource recycling and circular utilization, the precise classification of metal materials is an important link to achieve efficient resource recovery. However, due to the mixed distribution of different metals in waste, traditional sorting methods (such as those based on density, magnetism, or manual screening) are inefficient and have a high error rate, making it difficult to meet the needs of industrial development. In recent years, the analysis of spark flow characteristics has become an emerging method for metal classification. By using the differences in the spark morphology, color, trajectory, and dynamic characteristics generated by different metals during the grinding process, various metals can be effectively distinguished. However, the existing technology has difficulty meeting the real-time requirements in complex industrial scenarios. Especially in high-frequency metal recycling production lines, manual classification or the delay of existing classification systems will seriously affect the overall efficiency. Summary of the Invention
[0003] In order to solve the above problems, the purpose of the present invention is to provide an intelligent classification method for recycling and reusing metals based on dynamic analysis of spark flow videos, which realizes efficient online classification and resource sorting of metals in complex scenarios, effectively improves the efficiency of resource recycling, and reduces labor costs.
[0004] To achieve the above purpose, the present invention adopts the following technical solutions: An intelligent classification method for recycling and reusing metals based on dynamic analysis of spark flow videos, comprising the following steps: S1: Collect spark flow videos through a high-speed camera; S2: Decompose the spark flow videos into frame-by-frame images, select key frames, use a pre-trained YOLOv8 model to detect the spark regions in the videos, generate spark region segmentation results, and add median filtering and morphological operations to optimize the spark region segmentation results; S3: According to the segmentation results, store the information of each frame of the spark region as sequence data; S4: Extract the dynamic characteristics of the sparks based on the spark sequence data, associate the dynamic characteristics of the sparks with the time series, generate feature vectors, and construct a training dataset; S5: Construct a metal classification model based on a bidirectional LSTM model and train it based on the training dataset; S6: Deploy the trained metal classification model to a lightweight inference architecture, output the metal classification results according to the real-time video data of the sparks, and transmit the metal classification results to the garbage recycling system in real time.
[0005] Further, S1 is specifically as follows: Use a high-speed industrial camera with a frame rate set to 240 fps to ensure capturing the dynamic information of the rapidly spreading spark stream, and install a dust-proof glass and a stable light source filter to avoid the influence of dust and stray light in the industrial environment; Trigger the shooting through the force sensor on the grinding equipment and automatically record when sparks are generated during grinding; Collect spark stream videos under various metals, different grinding conditions, and background scenes.
[0006] Further, decompose the spark stream video into frame-by-frame images and select key frames, specifically as follows: Let the input spark stream video be V, with a total of T frames and a frame rate of f. Decompose the video into a frame sequence: {F1, F2, …, F t , …, F T}, and the t-th frame is the image F t ; Based on the dynamic characteristics of the sparks, select the frames with the strongest brightness and motion as key frames, and calculate the average brightness L(F t ) of each frame: ; where H and W are the height and width of the frame image, and F t (i, j) represents the pixel value, and i and j are the row and column indices of the pixel point; Select the frames with a brightness change greater than the threshold O :
[0007] where is the adjustment factor for the selection ratio, is the brightness threshold.
[0008] Further, use the deep learning model YOLOv8 to detect the spark area in the video, separate it from the background environment, generate a binary mask of the sparks, and add median filtering and morphological operations to optimize the spark area segmentation result: Use the pre-trained YOLOv8 model to detect the spark area in the frame image and generate the bounding box of the spark target: Given the input frame image F t , output the target box , where x and y are the center coordinates of the detection box, and w and h are the width and height; Use the box information B t to cut the spark area on the image: ; where represents a target boxes; M is the total number of target boxes; Grayscale the image and obtain the grayscale image G t Perform threshold segmentation on (i, j) in the grayscale image G to generate a binary mask M in the spark region t : ; where τ is the threshold value; Perform median filtering on the binary mask M of the spark region t to remove small-area noise and smooth the edges of the spark region: ; where MedianFilter represents the median filter and k is the filter window size; is the binary mask after median filtering; Increase the thickness of the boundary of the spark region through dilation operation to connect possibly disconnected particles: ; where K is the structuring element; ⊕ represents the dilation operation; is the binary mask output after the dilation operation; Finally, perform an erosion operation to remove noise points and retain the main spark region: ; where is the erosion operation; Result is the finally optimized spark region.
[0009] Furthermore, according to the segmentation result, store the information of each frame of the spark region as sequence data, specifically as follows: Based on the finally optimized spark region, find the minimum bounding rectangle of the spark region and calculate the maximum length L t : ; where is the set of contour points in the point pair; Count the number of pixels in the spark part of the binary mask and calculate the area A t : ; Calculate the centroid C of the spark region t : ; Organize the features of each frame of the spark region into a sequence to form the spark sequence data in the time dimension: .
[0010] Furthermore, dynamic features of the spark are extracted from the spark sequence data, including spark length features, spark particle movement trajectories, diffusion angles, and brightness features, as follows: For the finally optimized spark region , calculate the centroid C t to each edge point vector : ; Among them, is the coordinate of the edge point ; (x t , y t ) is the coordinate of the centroid C of the spark region at time step t t ; Calculate the angle between all vectors pairwise : ; Among them, is the vector of the edge point ; Take the maximum angle as the diffusion angle : ; For the original grayscale image G of the finally optimized spark region , calculate the average brightness of the spark region t : C t : ; Calculate the standard deviation of the spark brightness as the brightness feature, reflecting the volatility of the brightness: ; Associate the above features with the time series to generate the feature vector of each frame : ; Among them, is the total trajectory length, v t is the movement speed of the spark centroid; a t is the acceleration of the spark centroid.
[0011] Furthermore, a metal classification model is constructed based on the bidirectional LSTM model, as follows: Use bidirectional LSTM to process the time series features , capturing the forward and backward dependencies of the time series; ; Among them, is the forward memory unit at time t in the bidirectional LSTM; is the backward memory unit at time t in the bidirectional LSTM; is the forward hidden state at time t; is the backward hidden state at time t; LSTM fw is the forward LSTM; LSTM bw is the backward LSTM; The output of the bidirectional LSTM is the concatenation of the forward hidden state and the backward hidden state : h t = , ; An attention mechanism is added to the output of the bidirectional LSTM to dynamically select the most important frames in the time series for classification, and the attention weight α is calculated t : ; where tanh is the activation function; W a , b a are the weight and bias respectively; e t is the attention score at the current time t; The weighted sum is used to obtain the global feature F of the time series attn : ; The global feature F attn is input into the fully connected layer, and the classification result y is output: ; where Softmax is the activation function, W B is the weight of the fully connected layer, and b is the bias.
[0012] Furthermore, the training of the metal classification model is as follows: Based on the training feature dataset, the model is trained using the cross-entropy loss function to optimize the model:
[0013] where, is the number of metal categories, is the true label of the th sample; is the probability that the th sample predicted by the model belongs to the th class, and N' is the total number of samples; The Adam optimizer is used and a learning rate decay strategy is added: ; where λ is the decay rate; α0 is the learning rate; Dropout is added to the LSTM layer and the fully connected layer to prevent overfitting, and a weight regularization term is added to the loss function.
[0014] Furthermore, S6 is specifically as follows: Deploy the metal classification model on the TensorRT inference architecture to analyze the metal spark stream and output the classification result; Use the communication protocol to transmit the classification result to the resource sorting system in real time; After receiving the classification signal, the central control system converts the classification information into an execution instruction to interface with the robotic arm or sorting device; According to the transmitted metal classification result, guide the robotic arm or sorting equipment to divert different types of metals to complete resource classification.
[0015] The present invention has the following beneficial effects: 1. By the method of dynamically analyzing and classifying the spark stream through video, the present invention can significantly improve the classification efficiency of different materials in waste metals, transform the traditional method based on physical and chemical analysis into a non-contact, efficient and real-time intelligent classification method, reduce the recycling cost, and adapt to complex industrial scenarios; 2. The present invention uses the YOLOv8 model to detect the spark area, which can quickly identify the spark area and generate a segmentation result. Median filtering and morphological operations are used to optimize the segmentation result, remove noise and artifacts, ensure the clear boundary of the spark area, improve the robustness of the spark area detection, and provide higher-quality input for subsequent dynamic feature extraction; 3. By decomposing the spark stream video frame by frame and extracting the dynamic features of the spark (such as brightness, diffusion angle, centroid movement speed, trajectory length, etc.), the present invention can comprehensively describe the physical and dynamic characteristics of the spark stream, associate the spark area information with the time series, capture the dynamic law of the spark changing with time, and make up for the limitations of the single-frame image information. Description of the Drawings
[0016] Figure 1 is the flowchart of the method of the present invention. Detailed Embodiments
[0017] The following further describes the present invention in detail with reference to the drawings and specific embodiments: Refer to Figure 1 , in this embodiment, a method for intelligent classification of recycled metals based on dynamic analysis of spark stream video is provided, including the following steps: S1: Collect the spark stream video through a high-speed camera; S2: Decompose the spark stream video into frame-by-frame images, select key frames, use the pre-trained YOLOv8 model to detect the spark regions in the video, generate the spark region segmentation results, and add median filtering and morphological operations to optimize the spark region segmentation results; S3: According to the segmentation results, store the information of the spark region in each frame as sequence data; S4: Extract the dynamic features of the sparks based on the spark sequence data, associate the dynamic features of the sparks with the time series, generate feature vectors, and construct a training dataset; S5: Build a metal classification model based on the bidirectional LSTM model and train it based on the training dataset; S6: Deploy the trained metal classification model to a lightweight inference architecture to achieve real-time inference. According to the real-time video data of the sparks, output the metal classification results; and send the metal classification results to the garbage collection system in real time to guide the robotic arm or sorting device to perform resource classification, so as to achieve online classification and resource sorting.
[0018] In this embodiment, S1 is specifically: Use a high-speed industrial camera with a frame rate set to 240 fps or higher to ensure capturing the dynamic information of the rapid diffusion of the spark stream, and install a dust-proof glass and a stable light source filter to avoid the influence of dust and stray light in the industrial environment; Trigger the shooting through the force sensor on the grinding device and automatically record when sparks are generated during grinding; Collect spark stream videos under various metals (such as iron, aluminum, stainless steel), different grinding conditions (force, angle, rotation speed, etc.) and background scenes.
[0019] In this embodiment, decomposing the spark stream video into frame-by-frame images and selecting key frames are specifically as follows: Let the input spark stream video be V, with a total of T frames and a frame rate of f. Decompose the video into a frame sequence: {F1, F2, …, F t , …, F T}, and the t-th frame is the image F t ; Based on the dynamic characteristics of the sparks, select the frames with the strongest brightness and motion as key frames, and calculate the average brightness L(F t ) of each frame: ; where H and W are the height and width of the frame image, represents the pixel value, and i and j are the row and column indices of the pixel points; Select the frames with a brightness change greater than the threshold O (such as frames higher than twice the global average brightness):
[0020] Among them, is the adjustment factor for the selection ratio, is the brightness threshold.
[0021] In this embodiment, the deep learning model YOLOv8 is used to detect the spark area in the video, separate it from the background environment, generate a binary mask of the spark, and add median filtering and morphological operations to further optimize the spark area segmentation result: Use the pre-trained YOLOv8 model to detect the spark area in the frame image and generate the bounding box of the spark target: Given the input frame image F t , output the target box , where x and y are the center coordinates of the detection box, and w and h are the width and height; use the box information B t Cut the spark area on the image: ; Among them, represents a target boxes; M is the total number of target boxes; Grayscale the image, perform threshold segmentation on the grayscale image G t (i, j) to generate a binary mask M in the spark area t : ; Among them, τ is the threshold; Perform median filtering on the binary mask M of the spark area t to remove small-area noise and smooth the edge of the spark area: ; Among them, MedianFilter represents the median filter, and k is the filter window size; is the binary mask after median filtering; Through the dilation operation, increase the thickness of the boundary of the spark area and connect the possibly disconnected particles: ; Among them, K is the structuring element; ⊕ represents the dilation operation; is the binary mask output after the dilation operation; Finally, use the erosion operation to remove the noise points and retain the main spark area: ; Among them, is the erosion operation; The result is the finally optimized spark area.
[0022] In this embodiment, according to the segmentation result, the information of the spark region in each frame is stored as sequence data, specifically as follows: Based on the finally optimized spark region, find the minimum bounding rectangle of the spark region and calculate the maximum length L t : ; Among them, is a point pair in the contour point set ; Count the number of pixels in the spark part of the binary mask and calculate the area A t : ; Calculate the centroid C of the spark region t : ; Organize the features of the spark region in each frame into a sequence to form the spark sequence data in the time dimension: .
[0023] In this embodiment, dynamic features of the spark are extracted according to the spark sequence data, including spark length feature, spark particle movement trajectory, diffusion angle and brightness feature, specifically as follows: For the finally optimized spark region , calculate the centroid C t to each edge point vector : ; Among them, is the coordinate of the edge point ; (x t , y t ) is the coordinate of the centroid C of the spark region at time step t t ; Calculate the angle between all pairs of vectors : ; Among them, is the vector of the edge point ; Take the maximum angle as the diffusion angle : ; For the original grayscale image G of the finally optimized spark region t , calculate the average brightness of the spark region C t : ; Calculate the standard deviation of the spark brightness As the brightness feature, it reflects the volatility of the brightness: ; Associate the above features with the time series to generate a feature vector for each frame : ; Among them, is the total length of the trajectory, and v t is the moving speed of the spark centroid; a t is the acceleration of the spark centroid.
[0024] In this embodiment, a metal classification model is constructed based on a bidirectional LSTM model, specifically as follows: Use bidirectional LSTM to process time series features , capturing the forward and backward dependencies of the time series; ; Among them, is the forward memory unit at time t in the bidirectional LSTM; is the backward memory unit at time t in the bidirectional LSTM; is the forward hidden state at time t; is the backward hidden state at time t; LSTM fw is the forward LSTM; LSTM bw is the backward LSTM; The output of the bidirectional LSTM is the concatenation of the forward hidden state and the backward hidden state : h t = , ; Add an attention mechanism to the output of the bidirectional LSTM to dynamically select the most important frames in the time series for classification, and calculate the attention weight α t : ; Among them, tanh is the activation function; W a , b a are the weight and bias respectively; e t is the attention score at the current time t; Weighted summation to obtain the global feature F of the time series attn : ; The global feature Fattn Input the fully connected layer and output the classification result y: ; Among them, Softmax is the activation function, W B is the weight of the fully connected layer, and b is the bias.
[0025] In this embodiment, the metal classification model is trained as follows: Train based on the training feature dataset and optimize the model using the cross-entropy loss function:
[0026] Where, is the number of metal categories, is the true label of the th sample; is the probability that the th sample predicted by the model belongs to the th class, and N' is the total number of samples.
[0027] Use the Adam optimizer and add a learning rate decay strategy: ; Where λ is the decay rate; α0 is the learning rate; Add Dropout (such as 0.5) to the LSTM layer and the fully connected layer to prevent overfitting, and add a weight regularization term to the loss function.
[0028] In this embodiment, S6 is specifically: Deploy the metal classification model to the TensorRT inference architecture, analyze the metal spark stream, and output the classification result (such as "aluminum", "steel").
[0029] Use communication protocols (such as OPC UA, Modbus) to transmit the classification results (including metal categories and confidence levels) to the resource sorting system in real time; After receiving the classification signal, the central control system converts the classification information into an execution instruction to connect to the robotic arm or sorting device; According to the transmitted metal classification results, guide the robotic arm or sorting equipment to divert different types of metals to complete resource classification.
[0030] Those skilled in the art should understand that the embodiments of the present invention can be provided as a method, a system, or a computer program product. Therefore, the present invention can take the form of a complete hardware embodiment, a complete software embodiment, or an embodiment combining software and hardware aspects. Moreover, the present invention can take the form of a computer program product implemented on one or more computer-usable storage media (including but not limited to disk memory, CD-ROM, optical memory, etc.) that contain computer-usable program code.
[0031] The present invention is described with reference to the flowcharts and / or block diagrams of methods, apparatuses (systems), and computer program products according to embodiments of the present invention. It should be understood that each flow and / or block in the flowchart and / or block diagram, and the combination of flows and / or blocks in the flowchart and / or block diagram, can be implemented by computer program instructions. These computer program instructions can be provided to the processor of a general-purpose computer, a special-purpose computer, an embedded processor, or other programmable data processing devices to generate a machine, such that the instructions executed by the processor of the computer or other programmable data processing devices produce a means for implementing the functions specified in one Figure 1 one flow or multiple flows and / or blocks Figure 1 one block or multiple blocks.
[0032] These computer program instructions can also be stored in a computer-readable memory that can direct a computer or other programmable data processing devices to work in a specific manner, such that the instructions stored in the computer-readable memory produce a manufactured article including an instruction means that implements the functions specified in one Figure 1 one flow or multiple flows and / or blocks Figure 1 one block or multiple blocks.
[0033] These computer program instructions can also be loaded onto a computer or other programmable data processing devices, such that a series of operation steps are executed on the computer or other programmable devices to produce a computer-implemented process, and thus the instructions executed on the computer or other programmable devices provide steps for implementing the functions specified in one Figure 1 one flow or multiple flows and / or blocks Figure 1 one block or multiple blocks.
[0034] As described above, it is only the preferred embodiments of the present invention, and the present invention is not limited to other forms. Any person skilled in the art may use the technical content disclosed above to make changes or modifications into equivalent embodiments with equivalent changes. However, any simple modifications, equivalent changes, and modifications made to the above embodiments based on the technical essence of the present invention without departing from the technical solution content of the present invention still fall within the protection scope of the technical solution of the present invention.
Claims
1. An intelligent classification method for recycled metals based on dynamic analysis of spark stream videos, characterized in that, It includes the following steps: S1: Collect the spark stream video through a high-speed camera; S2: Decompose the spark stream video into frame-by-frame images, select key frames, use the pre-trained YOLOv8 model to detect the spark area in the video, generate the spark area segmentation result, and add median filtering and morphological operations to optimize the spark area segmentation result; S3: According to the segmentation result, store the information of each frame of the spark area as sequence data; S4: Extract the dynamic features of the spark according to the spark sequence data, associate the dynamic features of the spark with the time series, generate feature vectors, and construct a training data set; S5: Build a metal classification model based on the bidirectional LSTM model and train it based on the training data set; S6: Deploy the trained metal classification model to a lightweight inference architecture, output the metal classification result according to the real-time video data of the spark, and send the metal classification result to the garbage collection system in real time.
2. The intelligent classification method for recycled metals based on dynamic analysis of spark stream videos according to claim 1, characterized in that, The specific content of S1 is as follows: Use a high-speed industrial camera with a frame rate set to 240 fps to ensure capturing the dynamic information of the rapid diffusion of the spark stream, and install a dust-proof glass and a stable light source filter to avoid the influence of dust and stray light in the industrial environment; Trigger the shooting through the force sensor on the grinding device and automatically record when sparks are generated during grinding; Collect the spark stream video under various metals, different grinding conditions, and background scenes.
3. The intelligent classification method for recycled metals based on dynamic analysis of spark stream videos according to claim 1, characterized in that, The decomposition of the spark stream video into frame-by-frame images and the selection of key frames are specifically as follows: Let the input spark stream video be V, with a total of T frames and a frame rate of f. The video is disassembled into a frame sequence: {F1, F2, …, F t , …, F T}, and the t-th frame is the image F t ; Based on the dynamic characteristics of the spark, select the frame with the strongest brightness and motion as the key frame, and calculate the average brightness L(F t ): ; where H and W are the height and width of the frame image, and F t (i, j) represents the pixel value, and i and j are the row and column indices of the pixel point; Select frames with brightness changes greater than the threshold O : ; Among them, is the adjustment factor of the selection ratio, is the brightness threshold.
4. The intelligent classification method for recycled metals based on dynamic analysis of spark stream videos according to claim 3, characterized in that, Use the deep learning model YOLOv8 to detect the spark area in the video, separate it from the background environment, generate a binary mask of the spark, and add median filtering and morphological operations to optimize the spark area segmentation result: Use the pre-trained YOLOv8 model to detect the spark area in the frame image and generate the bounding box of the spark target: Given the input frame image F t , output the target bounding box , where x and y are the center coordinates of the detection box, and w and h are the width and height; use the box information B t Cut the spark area on the image: ; Among them, represents a target boxes; M is the total number of target boxes; Gray-scale the image to obtain the grayscale image G t Perform threshold segmentation on (i, j) to generate a binary mask M in the spark region t : ; where τ is the threshold; Median filter the binary mask M of the spark region t to remove small-area noise and smooth the edges of the spark region: ; Among them, MedianFilter represents the median filter, and k is the filter window size; is the binary mask after median filtering; Through the dilation operation, increase the thickness of the spark area boundary and connect the possibly disconnected particles: ; where K is a structuring element; ⊕ represents the dilation operation; is the binary mask output after the dilation operation; Finally, use the erosion operation to remove the noise points and retain the main spark area: ; Among them, is an etching operation; Result is the finally optimized spark region.
5. The intelligent classification method for recycled metal based on dynamic analysis of spark stream video according to claim 4, wherein, The storage of the information of each frame of the spark area as sequence data according to the segmentation result is specifically as follows: According to the finally optimized spark region, find the minimum bounding rectangle of the spark region and calculate the maximum length L t : ; Among them, is a set of contour points in the point pairs; Count the number of pixels in the spark part of the binary mask and calculate the area A t : ; Calculate the centroid C of the spark region t : ; Organize the features of each frame of the spark area into a sequence to form the spark sequence data in the time dimension: 。 6. The intelligent classification method for recycled metal based on dynamic analysis of spark stream video according to claim 1, wherein, The extraction of the dynamic features of the spark according to the spark sequence data includes the spark length feature, the movement trajectory of spark particles, the diffusion angle, and the brightness feature, specifically as follows: For the finally optimized spark region , calculate the centroid C t to each edge point vector : ; Among them, is the coordinate of the edge point ; (x t , y t ) is the coordinate of the centroid C t of the spark region at time step t; Calculate the angle between all pairs of vectors : ; Among them, is the vector of the edge point ; Take the maximum included angle as the diffusion angle : ; For the finally optimized spark region of the original grayscale image G t , calculate the average brightness of the spark region C t : ; Calculate the standard deviation of the spark brightness As a brightness feature, it reflects the volatility of the brightness: ; Associate the above features with the time series to generate a feature vector for each frame : ; Among them, is the total track length, and v t is the moving speed of the spark centroid; a t is the acceleration of the spark centroid.
7. The intelligent classification method for recycled metal based on dynamic analysis of spark stream video according to claim 6, wherein, The construction of the metal classification model based on the bidirectional LSTM model is specifically as follows: Using bidirectional LSTM to process time series features , capturing the forward and backward dependencies of the time series; ; Among them, is the forward memory cell at time t in the bidirectional LSTM; is the backward memory cell at time t in the bidirectional LSTM; is the forward hidden state at time t; is the backward hidden state at time t; LSTM fw is the forward LSTM; LSTM bw is the backward LSTM; The output of the bidirectional LSTM is the concatenation of the forward hidden state and the backward hidden state : h t =[ , ]; Add an attention mechanism to the output of the bidirectional LSTM to dynamically select the frames in the time series that are most important for classification, and calculate the attention weight α t : ; where tanh is the activation function; W a , b a are the weight and bias respectively; e t is the attention score at the current time t; The global feature F of the time series is obtained by weighted summation attn : ; Input the global feature F attn into the fully-connected layer and output the classification result y: ; Among them, Softmax is the activation function, W B is the weight of the fully connected layer, and b is the bias.
8. The intelligent classification method for recycling and reusing metals based on dynamic analysis of spark stream video according to claim 7, characterized in that, The training of the metal classification model is specifically as follows: Train based on the training feature data set and optimize the model using the cross-entropy loss function: ; Among them, is the number of metal categories, is the true label of the th sample; is the probability that the th sample predicted by the model belongs to the th class, and N' is the total number of samples; Use the Adam optimizer and add a learning rate decay strategy: ; where λ is the decay rate; α0 is the learning rate; Add Dropout to the LSTM layer and the fully connected layer to prevent overfitting, and add a weight regularization term to the loss function.
9. The intelligent classification method for recycling and reusing metals based on dynamic analysis of spark stream video according to claim 1, characterized in that, The specific content of S6 is as follows: Deploy the metal classification model to the TensorRT inference architecture, analyze the metal spark stream, and output the classification result; Use the communication protocol to transmit the classification result to the resource sorting system in real time; After receiving the classification signal, the central control system converts the classification information into execution instructions to interface with the robotic arm or sorting device; Based on the transmitted metal classification results, the robotic arm or sorting equipment is guided to divert different types of metals, completing resource classification.
Citation Information
Patent Citations
Yolov7 pantograph-catenary electric spark detection method based on attention mechanism
CN117789017A
Coal mine smoke and fire detection method based on improved YOLOv8 and related device
CN119919623A
Sorting of aluminum alloys
US20240246117A1
Sorting pieces of material based on photonic emissions resulting from multiple sources of stimuli
US7763820B1