Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

6464 results about "Image frame" patented technology

Skeleton detection and fall detection method based on improved spatio-temporal adaptive graph convolution

Disclosed in the present application is a skeleton detection and fall detection method based on improved spatio-temporal adaptive graph convolution. The method comprises the steps of: S1, collecting image data to acquire data of each image frame; S2: using a pre-trained yolov5 target person detection model to detect whether a target person is present in the data of each image frame, and if a target person is present, turning to step S3, and if no target person is present, ending the process; S3: for each detected target person, using a Deepsort target tracking algorithm to perform target tracking to obtain a tracking result, calculating the similarity to obtain the result of target association, and updating trajectory information of each target person; and S4: performing pose recognition on each target person on the basis of the trajectory information, using a spatio-temporal adaptive graph convolutional network to extract a feature vector of a pose, and using a classifier to perform human body behavior classification and recognition, in order to determine whether the target person has experienced a fall incident. The method achieves higher accuracy and robustness.
Owner:NANJING HOWSO TECH

Crop whole growth cycle identification method and system based on deep learning

The invention provides a crop whole growth cycle identification method and system based on deep learning, and the method comprises the steps: obtaining a multispectral image sequence of a target crop in a continuous time period through an image collection device, and carrying out the standardized illumination adjustment processing of each image frame in the multispectral image sequence, obtaining a standard illumination image set corresponding to the multispectral image sequence; carrying out crop region segmentation processing on each image, extracting a local feature region related to a target crop, generating a standardized crop image set containing the local feature region, inputting the standardized crop image set into a pre-trained multi-task deep learning model, extracting combined features through parallel convolution branches, and carrying out image segmentation processing on the combined features; and executing cross-stage correlation analysis in the full connection layer, outputting a multi-task classification result corresponding to the target crop growth cycle, and generating a stage identification report corresponding to the target crop full growth cycle. According to the invention, the crop whole growth cycle identification precision and the agricultural management efficiency can be improved.
Owner:HUAYUNSHENGDA(BEIJING)METEROLOGICAL TECH CO LTD

Self-adaptive camera pose intelligent adjusting system

The invention relates to the technical field of intelligent control, in particular to a self-adaptive camera pose intelligent adjusting system which comprises a dynamic recognition module, a pose adjusting module, a mechanism control module, an information filtering module and a feedback association module. According to the method, the foreground pixel brightness gradient and the boundary continuity in the image frame are extracted, the target displacement path can be recognized, the tracking information is constructed, the target coherent perception under rapid scene change is guaranteed, the attitude parameters and the target motion direction and speed are combined, the offset correction instruction is generated, the response stability of the dynamic target is improved, and the target tracking accuracy is improved. The execution structure feedback is used for correcting control output, reducing the driving error accumulation risk, eliminating non-target disturbance through image boundary definition and motion consistency, improving data purity and combining a view angle execution state and path information to realize synchronous control of tracking and view angle adjustment; and the response efficiency and the following robustness of the system in a dynamic environment are enhanced.
Owner:SHENZHEN STARCAM TECH

Three-dimensional dynamic scene reconstruction method and apparatus, and storage medium

The present disclosure relates to the field of computer vision and discloses a three-dimensional dynamic scene reconstruction method and apparatus, and a storage medium. The three-dimensional dynamic scene reconstruction method comprises: acquiring synchronized videos of a plurality of viewpoints of a dynamic scene; computing matching points between video images of different viewpoints, and estimating intrinsic and extrinsic parameters of each camera; obtaining a Gaussian splatting point set {p0} on the basis of a sparse point cloud constructed according to the depth of each matching point; for the first image frame of each video, using {p0} to perform static training thereon, to obtain a Gaussian splatting point set {p}; for the remaining image frames, dividing {p} into a static point set {S} and a dynamic point set {D}, performing dynamic training on {D}, and constructing a dynamic Gaussian splatting point set {P} from {p}, {S}, and the final {D}; and, in view of the intrinsic and extrinsic parameters of each camera, rendering {P} using a Gaussian splatting rendering pipeline, to obtain rendered images at different moments from new viewpoints.
Owner:TSINGHUA UNIVERSITY

Fire-fighting equipment detection early warning method and system based on visual camera

The invention belongs to the technical field of fire safety, and discloses a fire-fighting equipment detection early warning method and system based on a visual camera, and the method comprises the following steps: obtaining multi-focal-length image frames of a fixed spray head and a fire extinguisher edge region, judging the continuous change of gray scale in the continuous image frames, and carrying out the detection early warning of the fire-fighting equipment based on the visual camera. Comparing the gray level distribution difference under the long focus and the short focus, identifying the dynamic trajectory of the contour line angular point of the fire-fighting equipment association structure, analyzing the delay of the function action response, marking as an abnormal region, and generating a fire-fighting equipment collaborative early warning information set; according to the invention, by extracting and analyzing the multi-focal-length image frames of the edge areas of important facilities such as a fixed spray head and a fire extinguisher, fine changes of the fire-fighting facilities under differentiated focal lengths can be captured, working states and potential faults of the fire-fighting facilities can be judged, functional abnormalities caused by equipment aging or damage can be recognized in advance, early warning can be carried out in real time, and the safety of the fire-fighting facilities can be improved. And a higher guarantee level is brought to fire safety.
Owner:WEIFANG PING AN FIRE ENG CO LTD

Material intelligent transportation and safety monitoring system and method for shield construction

The invention relates to the technical field of tunnel engineering construction, and discloses an intelligent material transportation and safety monitoring system and method for shield construction, and the system comprises a visual perception unit, a sensor network module, an AI analysis center module, a safety decision module and a human-computer interaction interface. According to the invention, data acquisition is carried out through the visual perception unit and the sensor network module, multi-target detection operation is carried out on image frames through the AI analysis center module after target identification and track prediction, target types, space coordinates, contour boundaries and confidence coefficients are identified and extracted, and safety judgment and early warning output are carried out. According to the invention, by integrating the multi-view camera equipment and the UWB, GNSS and other sensors and adopting a deep learning target detection algorithm, high-precision identification and continuous tracking can be carried out on construction site personnel, equipment, segments and other key objects, and accurate input is provided for subsequent risk analysis.
Owner:CHINA RAILWAY 11TH BUREAU GRP CORP LTD +1

Network edge monitoring and early warning method based on video image AI analysis

The invention discloses a network edge monitoring and early warning method based on video image AI analysis, and the method comprises the following steps: S1, obtaining video data, processing the video data, and generating an image frame sequence; s2, analyzing an image frame sequence, and extracting space and time features; s3, taking the target as a hypergraph node, constructing a hyperedge based on space and time features, and dynamically adjusting hyperedge connection by using a genetic algorithm; s4, constructing a multi-layer hypergraph Transform network, extracting multi-scale spatial-temporal characteristics, and calculating a semantic relationship between nodes; s5, setting a butterfly optimization algorithm initial population, and dynamically optimizing model parameters through global and local search; s6, constructing an anomaly detection model, identifying an abnormal behavior, and feeding back a result to optimize model parameters and hyperedge selection; and S7, deploying the model at an edge node, triggering early warning when an abnormal behavior is detected, and pushing information to a management platform. According to the invention, through video image AI analysis, accurate detection and real-time early warning of abnormal behaviors in a network edge scene are realized.
Owner:SHAANXI VIDEO BIG DATA CONSTR & OPERATION CO LTD

Light guide plate defect detection method and system based on neural network

The invention discloses a light guide plate defect detection method and system based on a neural network, and particularly relates to the technical field of machine vision detection, and the method comprises the following steps: aiming at the problem of image instability of a light guide plate in a dynamic transmission or rotation process, continuously collecting an image sequence and extracting time domain features; and performing interference judgment in combination with the inter-frame consistency prediction coefficient and a first threshold to realize accurate identification of the abnormal image frame. For an abnormal image frame, further correcting the recognition credibility of the abnormal image frame by adopting a confidence adjustment and fusion mode, and meanwhile, introducing a frequency domain transformation and image enhancement strategy to compensate detail loss caused by motion blur; according to the method, inter-frame consistency analysis, confidence fusion regulation and control and frequency domain fuzzy recognition and compensation mechanisms are introduced, abnormal judgment and image quality restoration of the light guide plate image in the dynamic scene are realized, the recognition accuracy and stability of the neural network model on the defect type, position and confidence are improved, and the false detection and omission ratio is effectively reduced.
Owner:深圳市鸿卓电子有限公司

Coal mining equipment obstacle avoidance method and system based on machine vision

The invention relates to the field of excavation equipment obstacle avoidance based on image analysis, in particular to a coal mine excavation equipment obstacle avoidance method and system based on machine vision, and the method comprises the steps: obtaining the frequency domain characteristic parameters of the current electromagnetic noise during the operation of underground electromechanical equipment, and synchronously collecting the current original image frame output by a visual sensor; after a current de-noised image frame is obtained after dynamic de-noising processing, an obstacle contour feature point set is extracted through multi-scale edge detection, and an obstacle position confidence map is obtained according to space-time consistency analysis; obtaining point cloud coordinates of an image data loss area caused by electromagnetic interference, and mapping the point cloud coordinates to an image coordinate system to obtain a three-dimensional geometric complemented image; and constructing a dynamic occupation grid map of the underground environment according to the three-dimensional geometric complementation image and the obstacle position confidence map, and planning an obstacle avoidance path of the mining equipment according to the dynamic occupation grid map. And the obstacle avoidance capability of the mining equipment is improved, and safe and efficient coal mining operation is guaranteed.
Owner:TIANCHEN COAL MINE OF ZAOZHUANG MINING GRP

Concrete crack three-dimensional reconstruction method and system based on multi-modal fusion and medium

The invention discloses a concrete crack three-dimensional reconstruction method and system based on multi-modal fusion and a medium, and relates to the technical field of structural engineering detection, and the method comprises the following steps: obtaining a structural image frame sequence and a structural laser radar point cloud frame sequence of a target structure; obtaining a crack mask image frame after crack segmentation; an integral structure de-noised point cloud picture is obtained; obtaining a visible point cloud coloring point cloud map and a visible point cloud semantic segmentation point cloud map; obtaining an overall three-dimensional coloring point cloud map and an overall three-dimensional semantic segmentation point cloud map; obtaining an overall three-dimensional point cloud marking map; and performing three-dimensional attribute measurement on the crack based on the three-dimensional point cloud marking map to obtain three-dimensional geometric information of the crack. According to the method provided by the invention, a multi-frame and multi-modal fused crack structure reconstruction framework is designed, the method can adapt to crack detection of various three-dimensional structures, and simultaneous detection of crack width information, crack position and crack trend information can be realized based on an overall three-dimensional point cloud marking map.
Owner:CENT SOUTH UNIV

Space interaction accurate identification method and system based on multi-modal fusion

The invention provides a space interaction accurate recognition method and system based on multi-modal fusion, and relates to the technical field of artificial intelligence, and the method comprises the steps: collecting human skeleton, visual image and voice information, generating a space-time attention map through employing a feature extraction network, enhancing visual features, and executing feature complementary correction. Fusing the three types of modal information to obtain unified feature representation; when the recognition confidence is low, the space-time convolution generative adversarial network guided based on the action causal relationship graph complements the missing image frame, and the accuracy and robustness of space interaction recognition are improved.
Owner:ZHONGTIAN ZHILING (BEIJING) TECH CO LTD

Leakage detection and partial discharge digital imaging detection method and system based on acousto-optic fusion

The invention relates to the technical field of nondestructive testing, in particular to a leak detection and partial discharge digital imaging detection method and system based on acousto-optic fusion, and the method comprises the following steps: based on channel microphone array sound wave data in a partial discharge signal suspicious region, extracting a sound wave abnormal section, positioning a sound source, matching image edge features, and synchronously marking; and analyzing the phase change of the multi-frequency signal to judge a sound source concentration area, tracking the moving trend of a disturbance point, and outputting an acousto-optic fusion positioning trend track. According to the method, the partial discharge feature recognition sensitivity is improved through high-frequency peak paragraph screening and dominant frequency recognition, the abnormal region positioning precision is enhanced in combination with image edge extraction and sound source space matching, and acousto-optic synchronous positioning and trend trajectory display are achieved through multi-frequency signal phase analysis and image frame disturbance tracking. Through fusion of frequency domain feature extraction, image recognition, dynamic comparison and other actions, the spatial precision of abnormal source recognition and the multi-source fusion analysis efficiency are improved, and the partial discharge traceability and dynamic monitoring capability are enhanced.
Owner:李美娟 +1

Self-adaptive image stabilization and dynamic horizontal reference calibration method, device, equipment and medium

The invention relates to the technical field of image steady state adjustment, and provides a self-adaptive image stabilization and dynamic horizontal reference calibration method and device, equipment and a medium. Time domain alignment and motion parameter interpolation processing are carried out on original sensor data to obtain a synchronized motion parameter data stream, and multi-sensor data fusion filtering is carried out on the synchronized motion parameter data stream to obtain equipment three-dimensional attitude parameters. The method comprises the following steps: performing motion mapping on an original image frame according to equipment three-dimensional attitude parameters to obtain an image displacement compensation matrix, performing feature tracking and pixel resampling on the original image frame according to the image displacement compensation matrix to obtain a preliminary stable image set, and performing horizon detection and rotation compensation on the preliminary stable image set to obtain a horizontal reference alignment image set; and performing time sequence weighted fusion on the horizontal reference alignment image set to obtain a steady-state video data stream. Through multi-sensor fusion, high-precision mapping, image adaptive calibration and weighted fusion, the precision of motion jitter suppression and horizontal calibration is improved.
Owner:SHENZHEN WEIZHU TECH CO LTD

Self-adaptive deep fake face detection method and system based on space-frequency domain graph learning

The invention discloses a space-frequency domain graph learning-based adaptive deep fake face detection method and system, and the method comprises the steps: randomly extracting an image frame from a video, intercepting a face image, adjusting the feature dimension of the face image, and transmitting the face image to a depth adaptive wavelet module and a normalized residual homomorphic composition neural network module; a depth adaptive wavelet module extracts frequency features of the face image; a normalized residual homograph neural network module extracts spatial domain features of the face image; performing weighted fusion on the frequency domain features and the spatial domain features by using a self-adaptive feature fusion module based on gated convolution, realizing class attention guidance by using the gated convolution, dynamically adjusting the channel of a feature map and the weight of a spatial dimension, and finally obtaining fusion features; and performing classification according to the fusion features by using a classifier. According to the method, the extraction capability of forged detail clues is enhanced, and the detection precision and stability of the model are remarkably improved.
Owner:EAST CHINA JIAOTONG UNIVERSITY

Action localization method, device, electronic equipment, and computer-readable storage medium

An action localization method, device, electronic equipment, and computer-readable storage medium are provided. The action localization method includes: identifying at least one target video segment containing a target object in a video; acquiring a first action recognition result of at least one image frame in the at least one target video segment and a second action recognition result of the target video segment; and acquiring an action localization result of the video based on the first action recognition result and the second action recognition result.
Owner:SAMSUNG ELECTRONICS CO LTD

Intelligent gas hidden danger detection system

The invention relates to the technical field of anomaly detection, in particular to an intelligent gas hidden danger detection system which comprises a leakage monitoring module, a time difference extraction module, a direction positioning module, a color temperature analysis module and a risk aggregation module. According to the method, the time and space relation of gas concentration response nodes is extracted, a leakage diffusion path is constructed, and a dominant propagation direction is identified, so that a leakage trend region has a clear positioning basis, regional brightness and hue changes in an image frame are combined, the time sequence stability proportion is counted, and a cross-modal trend verification channel is established; the defect of a single sensing dimension is effectively made up, in the image and trend direction matching process, abnormal area numbers are recognized through stable proportion difference value changes, intersection confirmation is carried out on trend areas, the accuracy and spatial directivity of hidden danger marking are remarkably improved, the risk recognition chain is more complete through cooperative calculation of multi-source data, and the accuracy of hidden danger marking is improved. And the method has higher linkage response capability and dynamic judgment depth.
Owner:SHENZHEN ZHONGZHIAN QUALITY SAFETY TECH ASSESSMENT CENT CO LTD

Wind power construction intelligent safety management method and system based on intelligent AI monitoring

The invention relates to the field of image recognition, in particular to a wind power construction intelligent safety management method and system based on intelligent AI monitoring. The method comprises the following steps: obtaining an omnibearing real-time image flow of a wind power construction area, carrying out super-resolution deep convolution optimization and operator three-dimensional image segmentation, and extracting an operator three-dimensional image frame; three-dimensional point cloud modeling of the construction area is carried out based on the image flow, real-time image frame position positioning rendering is carried out according to a three-dimensional image frame, and a real-time twinborn model of the construction area is constructed; performing operation dynamic behavior analysis and behavior deviation degree quantitative analysis based on a twin model to obtain the behavior deviation degree of the operator; and according to the behavior deviation degree, carrying out early prediction analysis on illegal behaviors, making an adaptive risk early warning decision, and constructing an operation behavior risk early warning strategy. According to the invention, through real-time operation behavior identification and environmental risk analysis, the intelligence and safety level of wind power construction are improved.
Owner:JIANGXI QIANPING MASCH CO LTD

Pet target detection method and device and camera

The invention relates to the technical field of target detection, and discloses a pet target detection method and device and a camera, and the method comprises the steps: carrying out the motion triggering collection and image enhancement preprocessing of a front end region of a feeder, and obtaining an enhanced image frame sequence; performing feature extraction of dynamic receptive field adjustment on the enhanced image frame sequence to obtain pet feature descriptors and position information; behavior time sequence feature analysis is executed, and pet behavior sequence feature vectors are obtained; constructing a state transition diagram according to the pet behavior sequence feature vector, and performing time sequence consistency analysis to obtain a pet state judgment result; power management and decision execution are carried out on the feeder based on the pet state judgment result, feeding control under the low-power-consumption condition is achieved, behavior misjudgment caused by posture fluctuation is effectively avoided, the behavior recognition accuracy is improved, the accurate feeding control problem in a multi-pet family is solved, and the user experience is improved.
Owner:SHENZHEN ANKED SHITONG ELECTRONICS CO LTD

Unmanned aerial vehicle multichannel image transmission optimization system based on link quality perception

The invention provides an unmanned aerial vehicle multi-channel image transmission optimization system based on link quality perception so as to improve the transmission stability and the image quality guarantee capability of image data in a complex wireless environment. The prediction module constructs a time sequence model based on the historical link quality index of the wireless channel, and outputs a future link stability prediction value; the modeling module analyzes texture change and the like between the image frames, generates an image frame evolution vector, and calculates the aging sensitivity and reconstruction importance score of the image frames according to the image frame evolution vector; the decision-making module fuses the information and generates an optimal matching relation between the image frame and the channel and a transmission priority parameter; the grouping module performs image frame clustering according to the similarity between the priority parameters and the evolution vectors, constructs compression groups and generates corresponding compression configuration files and compression image data; the transmission module schedules the compressed image data to a corresponding wireless channel for transmission; the system can be widely applied to an unmanned aerial vehicle image transmission task with relatively high requirements on real-time performance and image quality in a dynamic environment.
Owner:SHENZHEN RUIWO MOBILE CO LTD

Defoaming agent foam distribution analysis method based on image feature recognition

The invention discloses a defoaming agent foam distribution analysis method based on image feature recognition, and particularly relates to the field of industrial foam behavior perception and analysis for recognizing an image object with a random mode as a feature, and the method comprises the following steps: obtaining a foam image sequence in a target area, and collecting the foam image sequence through imaging equipment, the image frames of the foam image sequence have time continuity; and performing disturbance feature extraction operation on the foam image sequence to obtain local disturbance speed information, membrane surface tension change trend information and form boundary fluctuation information of the foam edge within a preset time. According to the method, foam structure disturbance characteristics are extracted from an image time sequence, a structure evolution graph memory bank is constructed, and irregular sudden change image behaviors are identified in combination with a trend matching mechanism, so that dynamic perception and abnormal response of a sudden foam state without prior support are realized, and the problem that a random mode foam state cannot be identified is solved.
Owner:HANGZHOU SERAPH TECH CO LTD

Control method of cable for charging unmanned ship based on visual identification

The invention provides an unmanned ship charging cable control method based on visual identification, and relates to the technical field of data processing, and the method comprises the steps: collecting continuous image frames of an unmanned ship charging area, identifying the spatial displacement and inclination angle change of an unmanned ship, analyzing the attitude offset of the unmanned ship caused by sea waves, calculating a water surface disturbance factor, and calculating the water surface disturbance factor. The method comprises the following steps: identifying the position of a charging interface, obtaining an initial positioning coordinate, carrying out prediction analysis on a disturbance trend in a preset time window, predicting a position change range of the charging interface, obtaining prediction coordinate data, planning a butt joint path of a cable and the charging interface, obtaining a compensation path sequence, and generating a cable propulsion instruction. And collecting an area image of the charging interface to obtain path tracking image data, judging whether the butt joint of the charging interface is successful or not, and if not, dynamically correcting the compensation path sequence. According to the invention, the cable is controlled through visual identification to charge the unmanned ship.
Owner:TIMES TIANHAI TECHNOLOGY CO LTD

Newborn health assessment method and system based on visual analysis

The invention relates to the technical field of health assessment, in particular to a newborn health assessment method and system based on visual analysis, and the method comprises the following steps: obtaining a newborn image frame sequence, extracting a hue curve, recognizing a variation region to generate a layer, and extracting a mutation region bitmap group in combination with a heart rate RR interval difference value; and analyzing an included angle mapping grid between the center of gravity of the pigment and the heart rate slope, superposing a respiratory rate curvature, performing co-occurrence clustering to form a trend graph block, and evaluating a risk distribution interface formed by joint mutation. According to the method, a dynamic coupling relation between an image and a physiological signal is established through slope direction angle change and included angle deviation, cross-dimension mapping of heart rate trend change and visual feature change is achieved, the heart rate, skin color tracks and curvature change of respiratory frequency are fused, a trend gathering map is constructed, and the dynamic coupling relation between the image and the physiological signal is obtained. The risk judgment and display are completed according to the joint grading of the block density, the time span and the parameter quantity in the atlas, and the health level fluctuation trend is visually presented in the risk evolution sorting.
Owner:THE AFFILIATED HOSPITAL OF SHANDONG UNIV OF TCM

Security and protection monitoring video real-time transmission method based on Internet of Things

The invention discloses a security and protection monitoring video real-time transmission method based on the Internet of Things, and relates to the technical field of security and protection monitoring video transmission, and the method specifically comprises the following steps: when a texture repetition region exists in an image frame, determining all macro blocks forming the texture repetition region in the image frame, and marking the macro blocks as texture repetition blocks; performing comprehensive analysis on each texture repetition block, and evaluating a candidate motion vector direction dispersion degree when an image frame in the security and protection monitoring video has a texture repetition area; based on the evaluation result, determining whether to classify each texture repetition block and whether to respectively match a corresponding motion vector selection strategy; and according to the matched motion vector selection strategy, respectively executing corresponding dynamic regulation and control operations. According to the method, the problem that the direction of the motion vector is discrete and uncontrollable in a texture repetition area is solved, dynamic regulation and control of a coding strategy and synchronous optimization of a decoding end are realized, and the image reconstruction precision and the intelligent analysis stability are improved.
Owner:ANHUI CHANGTIAN INFORMATION TECH CO LTD

Multi-station PCBA board detection method based on machine vision

The invention relates to the technical field of automatic detection, in particular to a multi-station PCBA (printed circuit board assembly) detection method based on machine vision, which comprises the following steps: acquiring a multi-station reference pulse period, dividing a time window, comparing starting time difference, controlling image acquisition to synchronize, locally imaging a PCBA based on a synchronizing signal, extracting temperature response of a welding spot, and constructing a heat conduction path. Mapping to a visible light image, evaluating welding spot symmetry, extracting welding spot hue change, marking abnormal frames, recording defect path nodes, and generating a defect path structure. According to the method, the temperature response coordinate points in the welding spot area are extracted and the heat conduction direction vector path is constructed, so that the welding quality is evaluated more carefully and accurately, the thermal grid chart is mapped to the visible light image frame, the image analysis process is further optimized, the defect detection accuracy is enhanced, and the potential defect area is effectively identified; the efficiency and reliability of PCBA board detection are significantly improved, and powerful technical support is provided for high-quality production.
Owner:广东德智矩阵科技有限公司

Multi-target tracking method based on YOLOv8 model and Byte Track algorithm

The invention provides a multi-target tracking method based on a YOLOv8 model and a Byte Track algorithm, and relates to the technical field of computer vision and edge equipment. The method specifically comprises the following steps: acquiring data of a plurality of images, performing format conversion, and constructing an image data set; constructing a DC-YOLOv8 network structure, and performing training by using the image data set to obtain a target detection model based on the DC-YOLOv8 network structure; obtaining a to-be-detected video stream, extracting continuous image frames from the to-be-detected video stream, and preprocessing the extracted image frames; and inputting the preprocessed image frames into the target detection model for target detection, and performing target tracking on all detected targets by adopting a Byte Track algorithm to obtain tracking trajectories of the tracked targets in all the image frames and generate a video stream. According to the invention, target tracking can be carried out on a video with many targets more smoothly.
Owner:NORTHEASTERN UNIV CHINA

Edge-deployed semi-supervised anomaly detection method and system for railway track foreign object

Disclosed in the present invention are an edge-deployed semi-supervised anomaly detection method and system for a railway track foreign object. The method comprises the following steps: an edge device encoding and decoding a video stream captured by a camera to obtain an image frame sequence, and performing frame extraction; and using a semantic segmentation model to perform image segmentation on a certain image frame obtained by means of frame extraction, to obtain a railway track region segmentation image. The use of a single image as input may generate an expert model result having a high weight value; however, the determination based on a single image is not stable, multiple consecutive images of the task scene need to be inputted, the frequency of each expert model obtaining the highest weight is computed, and the expert model corresponding to the highest frequency is the final solution. The present invention supports scene-adaptive foreign object detection algorithm automatic selection, and a user can perform selection on the basis of prior knowledge, or selection may be performed by a scene-adaptive automatic algorithm selection method; the user only needs to provide a batch of image data of the current scene, and the optimal algorithm selection can be evaluated.
Owner:GUANGZHOU EMBEDDED MACHINE TECH CO LTD

3D object detection using temporal inputs

Apparatuses, systems, and techniques of using one or more machine learning processes (e.g., neural network(s)) to detect objects from a plurality of image frames. In at least one embodiment, a plurality of image frames are fused into a feature map using one or more neural networks. In at least one embodiment, a plurality of image frames are processed using one or more neural networks to detect objects in a 3D space.
Owner:NVIDIA CORP

Hull surface defect detection system based on machine vision

The invention provides a hull surface defect detection system based on machine vision, and relates to the technical field of data processing. The image correction module is used for carrying out illumination equalization processing and geometric distortion correction; the region construction module is used for identifying a defect-free stable region and generating reference region data which comprises a brightness model and a texture model; the candidate generation module is used for detecting a region where texture interruption or abnormal bright spots exist locally to form candidate defect data, and the candidate defect data comprise pixel positions and local contrast parameters; the stability judgment module is used for carrying out projection matching in the multiple frames of images and simultaneously carrying out joint comparison with the brightness model and the texture model of the reference area data to form real defect data and false defect data; the result output module is used for generating a detection result containing defect coordinates, defect contours, image frame numbers and interference sample prompts; the accuracy of hull surface defect detection is improved.
Owner:福建博洋船舶工业有限公司

Methods and processors for rendering a 3D object using multi-camera image inputs

Methods and processors for rendering a 3D object are disclosed. The method includes acquiring multi-camera image input including first image frames of the 3D object generated by a first camera and second image frames of the 3D object generated by a second camera, acquiring an initial 3D Gaussian Splatting (3DGS) model having a plurality of initial parameters including an initial frame-wise GS parameter and an initial camera-wise GS parameter, generating an adjusted 3DGS model by adjusting, based on the multi-camera image input, at least one of: the initial frame-wise GS parameter, the initial camera-wise GS parameter, generating, by the adjusted 3DGS model, a 3DGS output and rendering a 2D image of the 3D object using the 3DGS output.
Owner:YINWANG INTELLIGENT TECHNOLOGIES CO LTD

Driver fatigue monitoring method based on eye movement tracking

The invention provides a driver fatigue monitoring technology using eye movement tracking, and relates to the field of auxiliary driving, and the technology comprises the steps: collecting a face image frame through a front-end camera, carrying out Medipe (Media Pipeline) detection, recording a head posture, and recognizing an iris center region; extracting continuous eye closing time and mouth state information of the driver through the face information; a PNP algorithm is used; a world coordinate system and a camera coordinate system are established, and the sight line vector and the external environment are located in the same space; modeling is carried out on the human eyes and the fixation area; coordinates of a fixation point in a three-dimensional space are obtained through calculation of the geometric model; real-time attention analysis is carried out through the fixation point, and an analysis result is determined; the attention change condition and the fatigue state of the driver are effectively recognized through a BP neural network in combination with a fixation point analysis result. And early warning measures are taken in time to remind a driver. The system is crucial for improving the driving safety, and necessary warning and intervention can be provided when the driver is distracted or fatigued.
Owner:HOHAI UNIV