Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

151 results about "Associated image" patented technology

Geosynchronization of an aerial image using localizing multiple features

A georegistration (a.k.a. georectification) of an image captured by a camera in an aerial vehicle, such as a satellite, is based on identifying multiple features using descriptor sets, and sending to a ground station only the descriptors of the identified features and the associated locations in the captured image, without sending of the captured image itself, thus requiring a low communication bandwidth. Using a database of geosynchronized reference images, the ground station uses the received descriptors sets and the associated image locations to localize the features on a selected geosynchronized reference image from the database, and forms a mapping function that map any locations in the captured image to geographical coordinates on Earth. The mapping may be used to geosynchronize an additional feature identified in the aerial vehicle, or to geo synchronize a region that may be cropped from the captured image and sent to the ground station.
Owner:EDGY BEES LTD

Monitoring video viewing method and device, computer equipment and medium

The invention relates to a monitoring video viewing method and device, computer equipment and a medium, and the method comprises the steps: determining a source camera of a target monitoring video needing to be viewed and corresponding viewing time information in response to a monitoring video viewing instruction; searching a camera to which an external monitoring video having a view overlapping area with the target monitoring video within the same time according to the viewing time information, and taking the camera as a part of neighbor cameras of the source camera; determining the optimal privacy metadata from the privacy metadata of the corresponding viewing time information of the partial neighbor cameras, wherein the privacy metadata comprises frame associated image blocks corresponding to the shielded areas of the privacy sensitive targets in the external monitoring videos of the corresponding neighbor cameras; and repairing a covered area of a corresponding privacy sensitive target in the target monitoring video according to the frame associated image block of the optimal privacy metadata, and then playing the target monitoring video. According to the invention, on the premise of default privacy protection, key details of the monitoring picture are intelligently restored and enhanced as required.
Owner:深圳市灵智无界科技有限公司

Applications for gain curves in imaging and video

Techniques are disclosed relating to exchange of images in networked computing applications. In particular, the disclosure relates to exchange of gain curves that are used to represent imaging and / or video in such applications. A gain curve may define a mathematical transformation that relates values from a source image domain to a destination image domain. The image and its associated gain curve(s) may be published to destination devices for consumption. When a destination device consumes the image, the destination device may apply a transform to source image content according to the gain curve(s) published with the image. For example, the destination device may apply a gain curve to an associated image directly, or it may derive another transform from the gain curve and additional information known to the destination device.
Owner:APPLE INC

Physician-guided machine learning system for assessing medical images to facilitate locating of a historical twin

A computer-implemented method of evaluating a user image of a patient to enable identification of a historical twin of the patient. The method includes organizing a plurality of medical images in an archive and receiving from a medical professional each of: (i) a region of interest; (ii) a textual description; (iii) selections for binary criteria; and (iv) weights of weighable criteria. The method comprises using a natural language search to create a relevant set of medical images and creating an optimal set from the relevant set of medical images by discarding medical images from the relevant set based at least on the selections for binary criteria. The method includes image processing medical images in the optimal set using the weight of the features of the region of interest to create medical image results. The relevant set comprises less than ten percent of the medical images in the archive.
Owner:IRANI NEVILLE

Region-text caption generation using global caption information

Approaches presented herein may be used to generate captions using raw caption information. Raw caption information may be used, with an associated image, to generate a detailed image caption. Object lists may then be generated from the image and / or the detailed image caption to produce an image including boxing box proposals for objects within the image. One or more trained machine learning systems may then be used to generate region of interest captions that infuse the global caption context associated with the raw caption information.
Owner:NVIDIA CORP

Aviation oil pipeline unmanned aerial vehicle intelligent inspection method and system

The invention discloses an aviation oil pipeline unmanned aerial vehicle intelligent inspection method and system, and the method comprises the steps: collecting visual image data along a pipeline through an inspection terminal carried by an unmanned aerial vehicle, covering a pipeline body and a surrounding environment, and recognizing an abnormal scene endangering the safety of the pipeline; inputting the image data into an abnormal scene recognition model deployed at an unmanned aerial vehicle end, and judging whether a preset type of abnormality is included; if at least one type of abnormity is identified, generating an alarm signal; alarm and related images are uploaded to a remote monitoring center server through wireless communication, and intelligent unmanned inspection of the running state of the pipeline is achieved. An image acquisition and recognition model is integrated at an unmanned aerial vehicle end, a pipeline and an environment are sensed in real time, and key abnormity is automatically recognized; the edge deployment model reduces invalid return, and only triggers alarm uploading when a risk is detected; and in combination with wireless communication return alarms, high-reliability monitoring is realized, and intelligent and refined guarantee is provided for safe operation of aviation oil pipelines.
Owner:CHINA AVIATION OIL PENGZHOU PIPELINE TRANSPORTATION CO LTD

Multi-modal map enhanced retrieval method and dialogue system based on feature fusion optimization

The invention discloses a feature fusion optimization-based multi-modal map enhancement retrieval method and a dialogue system. The method comprises the following steps of: respectively carrying out pre-training and fine tuning on a visual model and a language model by utilizing a domain image and text data; constructing a knowledge graph based on the text data in the knowledge base and constructing a vector database containing associated image data; performing semantic analysis and optimization on the original query of the user by using the language model and forming a structured retrieval intention; searching related sub-graphs, text semantic vector information and associated image data based on the search intention; encoding the sub-images into knowledge contexts, inputting the knowledge contexts into a dynamic prompt generator to generate visual prompts, and extracting enhanced visual features from the associated image data through a visual model; and inputting the subgraph, the text semantic vector information and the enhanced visual features into a language model for collaborative reasoning, and generating and outputting a final answer. According to the method, deep fusion and accurate retrieval of multi-modal knowledge can be realized, and the accuracy and efficiency are remarkably improved.
Owner:ZHEJIANG UNIV

Method, computer device, and computer-readable recording medium to provide message summary and associated image

A method of providing a message summary and an associated image may include requesting an image search in relation to a message summary created based on a message in a chatroom; receiving an image bundle that includes at least one image in response to an image search request; and displaying the image bundle in association with the message summary.
Owner:LINE PLUS

Power system-oriented dynamic knowledge base driven dialogue generation system, method, equipment and medium

The invention discloses a power system-oriented dynamic knowledge base driven dialogue generation system, method, equipment and medium, and the system comprises a dialect collection and recognition module which is used for collecting dialect voice data under different regional power scenes, carrying out the noise reduction of the dialect voice data, carrying out the dialect recognition through a deep learning model, and obtaining a dialect recognition result; converting the dialect voice data into a standard text; a dynamic knowledge base module; the language processing and image reasoning module is used for receiving the standard text, performing semantic understanding, generating semantic representation in combination with a knowledge base in the dynamic knowledge base module, and performing target detection and recognition on a power equipment fault related image input by a user to obtain an image recognition result; and a dialogue generation module. According to the invention, a cooperative system of four modules of dialect acquisition and identification, a dynamic knowledge base, language processing and image reasoning and dialogue generation is constructed, so that intelligent dialogue service oriented to the power industry is realized.
Owner:GUIZHOU POWER GRID CO LTD

system

A system is provided.SOLUTION: A system comprising: means for a user to upload an article; means for a server to parse the uploaded article; means for the server to automatically translate the parsed article into multiple languages; means for the server to extract and embed SEO keywords for the translated article; means for the server to generate and link relevant images and video to the article; and means for the server to distribute the processed article to platforms in each country.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

Poster background image selection, model training, poster generation method and related device

The application discloses a poster background image selection, model training, poster generation method and related device. Through the application of the technical solution, a plurality of weakly related image text pairs can be used to train a preset visual text model, and a visual text model obtained through the training is used to automatically select a poster background image weakly related to the text information of interest to a user. Then, a final poster image is generated based on the automatically selected poster background image. Thus, the problem that a large number of high-quality poster requirements cannot be met by relying only on manual design to generate posters in the related art is avoided.
Owner:BEIJING ACAD OF ARTIFICIAL INTELLLIGENCE +1

Three-dimensional reconstruction method and related apparatus

The embodiment of the application provides a three-dimensional reconstruction method and related devices. The method comprises: based on an image frame, extracting point data from point cloud data associated with the image frame and putting the point data into a target point data set; wherein, the associated image frame and the point cloud data are: there is an overlapping surface area between a surface area of a target object corresponding to the image frame and a surface area of the target object corresponding to the point cloud data; the point data put into the target point data set corresponds to the overlapping surface area; in the process of continuously scanning to generate the associated image frame and the point cloud data, point data meeting a specified condition in the point cloud data is added to the target point data set; and a three-dimensional reconstruction model of the target object is generated according to the point data in the target point data set. The three-dimensional reconstruction efficiency can be improved to a certain extent.
Owner:SCANTECH (HANGZHOU) CO LTD

Method for multispectral recording of an image stream and associated image recording system

A method for multispectral recording of an image stream, in which a sequence of single images of a scene, in particular a continuous video image data stream, is recorded as an image stream using an image sensor of an image recording system. At least two different types of single images are recorded in different associated wavelength ranges using the image sensor. At least one type A single image of the sequence is recorded during a chronological type A recording segment and at least one type B single image of the sequence is recorded during a chronological type B recording segment. The type A recording segment and the type B recording segment are chronologically separated from one another by a respective waiting interval, in which no image data are sensorially acquired using the image sensor. An image recording system is also provided that to carry out the method.
Owner:SCHOLLY FIBEROPTIC GMBH

Distributed camera system

A system having a central station and a plurality of cameras installed various locations. To search for and locate an item of interest, the central station generates and sends an item model to the cameras. When stored in a camera, the item model causes a logic circuit of the camera (e.g., a deep learning accelerator) to use image data, received from an image sensor for storing in a memory device of the camera, as an input to an artificial neural network. The logic circuit performs the matrix computation of the artificial neural network to generate a classification of whether the images are relevant to the item of interest characterized by the item model. If so, the camera transmits the relevant images to the central station for further processing to determine a real time location of the item of interest.
Owner:MICRON TECHNOLOGY INC

Audible auditing method and device, computer equipment, readable storage medium and program product

The invention relates to an added staff auditing method and device, computer equipment, a readable storage medium and a program product. The method comprises the following steps: when an auditing node is transferred to a first target node of a target workflow, acquiring an added personnel auditing associated image of a to-be-audited object, inputting the added personnel auditing associated image into an auditing model for extracting to-be-audited information to obtain auditing reference information, and finally, responding to an added personnel auditing operation for the to-be-audited object, and performing auditing on the to-be-audited object. And displaying the auditing reference information and the to-be-audited information of the to-be-audited object in a display interface, so as to determine an auditing result of the to-be-audited object at the first target node based on the auditing reference information and the to-be-audited information of the to-be-audited object. By adopting the method, the working efficiency of staff increase auditing can be remarkably improved.
Owner:CHINA LIFE INSURANCE CO LTD

Electric energy meter appearance defect detection method and system based on AI automatic identification

The invention relates to the technical field of electric energy meter appearance detection, and discloses an AI automatic identification-based electric energy meter appearance defect detection method and system. The method comprises the steps that electric energy meters are grouped to form detection units, and a detection device is started after an initial detection instruction is received; when initial detection is executed, basic appearance information of the electric energy meter of each detection unit is collected, and corresponding historical defect records are called; the defect probability grade and the criticality grade of each detection unit are evaluated by combining the two types of information, and a personalized detection sequence of the detection device is planned accordingly. The detection device captures real-time image information of the electric energy meter when executing tasks according to the sequence, obtains field interference factor data to judge whether the electric energy meter is interfered or not, and corrects the electric energy meter when the electric energy meter is interfered and a detection track deviates; the real-time image information is analyzed through an artificial intelligence algorithm to find appearance flaws, an alarm is generated after the appearance flaws are found, and flaw sites and related image information are transmitted to a user display end.
Owner:HANGZHOU WOYI DIGITAL TECH CO LTD

Video processing system for non-inductive stacking of objects by using video tiles

There is described a video processing system (10), in particular as part of a television relaying system, for processing video frames (VE, VEc) of a sequence of video images, in particular of a video stream, in which the video processing system (10) is adapted for providing, for each video frame (VE, VEc), at least one associated image mask (PM, VM) and associated camera data (KD), and wherein the video processing system (10) has: a video frame operation unit (12); a video tile generation unit (14); the video tile generation unit has:-a texture tile generation unit (16) adapted for calculating a texture tile (TP) based on a video frame (VE) and at least one image mask (PM, VM) associated with this video frame (VE); -a reference storage unit (18) adapted for storing at least two reference images (RB1, RB2) and respective reference texture tiles (RT1, RT2), where one reference image (RB1, RB2) corresponds to one video frame (VE) as a basis for the calculation of a texture tile (TP), where the texture tiles are stored as reference texture tiles (RT1, RT2); -an adjustment unit (20) adapted for calculating an associated video tile (VP) for each video frame (VEc) input to the video processing system (10), in which the stored at least two reference images (RB1, RB2) and reference texture tiles (RT1, RT2) are taken into account during the calculation.
Owner:UNIQFEED AG

Page display method and device, equipment, storage medium and product

PendingCN121455367ACustomer relationshipTransmissionComputer hardwareItem Collection
The invention provides a page display method and device, equipment, a storage medium and a product, and relates to the technical field of computers. The method comprises the steps that an information input item set is displayed in a first area in a preset page, preset media content is displayed in a second area in the preset page, and the preset media content comprises image information related to at least one information input item. According to the technical scheme, the information input item set and the image information related to the at least one information input item are displayed in the same page at the same time, the related image information of the information input item can be displayed more visually and vividly, a user is helped to know the information input item more comprehensively, and the user experience is improved. And the information input efficiency, accuracy and success rate are improved.
Owner:CHENGDU GUANGHEXINHAO TECH CO LTD

Visual retrieval augmented generation for multimodal large language models

Systems and methods for visual retrieval augmented generation for artificial intelligence models such as multimodal large language models. Associations between image and description pairs can be identified from an awareness dataset by finetuning a multi-modal large language model (MLLM) with the awareness dataset based on randomly chosen images added to each example from a relevant dataset. Visual distractions for image processing with the MLLM can be minimized by finetuning the MLLM with a focus dataset based on randomly chosen images added to each example from the relevant dataset. Visual hallucinations from the MLLM can be mitigated by finetuning the MLLM with a learning dataset based on related images having corresponding texts added to each example from the relevant dataset to utilize extracted information from associations between provided text from multiple images and a learning dataset.
Owner:NEC LABORATORIES AMERICA INC

Exposure convergence method and related image processing device

ActiveUS12604097B2Imaging processingRadiology
An exposure convergence method of showing forward change of image intensity in response to drastic change of ambient light and applied to an image processing device includes utilizing an exposure function to acquire an initial intensity of a detection image provided by an image sensor, providing an exposure setting parameter to the image sensor for generating an input image with a first intensity in accordance with an analysis result of the initial intensity, and utilizing a gain parameter to adjust the input image for generating an output image with a second intensity, so that the second intensity of a plurality of sequential output images is gradually varied in one direction.
Owner:MEDIATEK INC

Context-based image state selection

System, method, and non-transitory computer readable medium for presenting images on a mobile device. Images are presented by monitoring one or more physical characteristics surrounding the mobile device using at least one sensor, determining a contextual state of the mobile device based on the monitored one or more physical characteristics, selecting an image from a plurality of related images associated with the determined contextual state, generating at least one overlay image from the selected image, and presenting the at least one overlay image with an optical assembly of the mobile device.
Owner:SNAP INC

Question and answer processing method and electronic equipment

The invention provides a question and answer processing method and electronic equipment. The method comprises the following steps: receiving a question text and a question image input by a user; intention recognition is conducted on the question text, a task description text is generated, and the task description text comprises slot position information; vectorizing the task description text and the problem image to obtain a text vector and an image vector; based on a multi-modal database and a multi-modal index, recalling target text knowledge and associated image knowledge corresponding to the text vector, recalling target image knowledge and associated text knowledge corresponding to the image vector, and recalling key field information corresponding to slot information; optimizing the target text knowledge, the associated image knowledge, the target image knowledge, the associated text knowledge and the field information, and then performing aggregation to obtain an aggregation result; and based on the aggregation result, generating a question-answer result by using a large language model. According to the method, the accuracy and credibility of the question and answer result can be improved.
Owner:ZHEJIANG GEELY HLDG GRP CO LTD +1

Factory floor modeling

System and techniques for facility floor modeling are described herein. A scan set that includes depth data and correlated image data from various positions on a facility floor is obtained. This data includes representations of equipment in the factory. A trained model is invoked, receiving the scan set as input and producing, as output, a computer model of the equipment along with associated information. The computer model of the equipment is integrated into a larger computer model of the factory and the information regarding the equipment is stored in a data structure that corresponds to the equipment's representation within the computer model of the factory.
Owner:INTEL CORP

Smoke detection method, device and equipment based on multi-modal fusion

The invention relates to a smoke detection method, device and equipment based on multi-modal fusion. The method comprises the following steps: controlling an acquisition unit to acquire image data and point cloud data of a scene to be monitored through a synchronization unit mounted on a rail-mounted inspection robot; based on a pre-established initial pixel point cloud mapping relationship, performing spatial calibration and dynamic association on the image data and the point cloud data, and performing multilevel feature extraction on the image data and the point cloud data after spatial calibration and dynamic association to obtain an image feature map and a point cloud feature map; based on a preset cross-modal feature interaction mechanism, fusing the image feature map and the point cloud feature map to obtain a fused feature map; and performing smoke detection on the fused feature map and outputting a detection report. The problems that in the prior art, accurate space alignment and an effective feature interaction mechanism are lacked between point cloud and image data, feature information fusion is poor, and detection accuracy is low are solved.
Owner:GUANGZHOU GUOXUN ROBOT TECH CO LTD

Video coding and decoding

A sequence of images is encoded in a bitstream as a series of picture units PU-01˜03. Each picture unit corresponds to one encoded image and includes one or more network abstraction layer (NAL) units NAL-01˜23. The NAL units may be video coding layer (VCL) NAL units which each contain encoded image data or adaptation parameter set NAL units which each contain an adaptation parameter set (APS) having parameters for performing one or more types of processing operation on the image data contained in one or more VCL NAL units. The APS NAL units may be prefix APS NAL units P-APS or suffix APS NAL units S-APS. An additional constraint is applied to the bitstream prohibiting inclusion of a prefix APS NAL unit after the first NAL unit of the picture unit concerned. This can avoid more than one APS applying to slices belonging to the same picture unit.
Owner:CANON KK

Muscarinic poisoning identification method and system

The invention discloses a muscarinic intoxication identification method and system, which are applied to the technical field of artificial intelligence, and the method comprises the steps: obtaining a muscarinic data set; wherein the muscarinic data set comprises a plurality of case samples, each case sample comprises a muscarinic associated image, chief complaint text information and biochemical test information, and each case sample has a corresponding label; performing model training processing by using the muscarinic data set to obtain a muscarinic poisoning identification model; performing identification processing on the to-be-detected case sample by using the muscarinic intoxication identification model to obtain a muscarinic intoxication identification result; wherein the to-be-tested case sample comprises a to-be-tested muscarinic associated image, to-be-tested chief complaint text information and to-be-tested biochemical test information. According to the invention, the recognition precision of muscarinic poisoning can be effectively improved.
Owner:SHANTOU UNIV

A deep learning-based high-throughput visual analysis method and system for corn kernel and embryo volume

The present application relates to the technical field of image data processing, and discloses a corn kernel and embryo volume high-throughput visual analysis method and system based on deep learning, which can obtain at least one folder according to a file input path set by a user, and each folder includes a gray scale sequence obtained by computer tomography (CT) slicing of corn kernel groups in a horizontal direction. Based on a parallel pipeline processing mode, a set batch size and a deep learning segmentation model, the gray scale sequence in the folder is preprocessed and image segmented to obtain kernel binary segmentation mask sequences and embryo binary segmentation mask sequences, and then the kernel quantity and the volume related data of the kernel and the embryo are determined, and the related images and data are visually output. The present application can complete the processing of the high-throughput data of the CT slice sequence of the corn kernel, and the nondestructive measurement and visual display of the volume related data of the kernel and the embryo, and effectively improve the measurement efficiency.
Owner:CHINA AGRI UNIV

Security video stream real-time anomaly recognition method and system based on edge computing

PendingCN122368890AEdge nodeConfidence metric
This invention relates to the field of video recognition technology, and discloses a method and system for real-time anomaly recognition of security video streams based on edge computing. The method includes: acquiring real-time video streams from monitoring devices at edge nodes, decoding and standardizing them to obtain a sequence of image frames of the target area; extracting features from each frame of the sequence and performing preliminary anomaly detection to generate a preliminary anomaly detection result; when the result indicates the presence of a potential anomaly, triggering deep behavioral analysis, performing spatiotemporal feature fusion analysis on the associated continuous image frame sequence to obtain a behavioral semantic description vector of the target area; calculating the deviation of this vector from the normal behavioral pattern cluster at the edge nodes to determine the anomaly type label and confidence level; generating and issuing a graded alarm according to a preset alarm strategy, and simultaneously packaging the anomaly label, associated image frame summary, and alarm signal into a structured log and uploading it to a cloud data center. This invention can improve the efficiency of real-time anomaly recognition of security video streams.
Owner:BEIJING GADE WEILAI TECHNOLOGY CO LTD

Data processing method and device, analysis system, electronic equipment and storage medium

The invention discloses a data processing method and device, an analysis system, electronic equipment and a storage medium. The data processing method comprises the following steps: based on an association relationship between a plurality of image acquisition devices and a plurality of identification algorithms, performing identification processing on an image acquired by the associated image acquisition device by using each identification algorithm to obtain a plurality of groups of identification data corresponding to the plurality of identification algorithms, configuration information is set according to the incidence relation between the multiple image acquisition devices and the multiple recognition algorithms; and based on the plurality of statistical rules, performing statistical analysis on the plurality of groups of identification data to obtain a plurality of statistical data corresponding to the plurality of statistical rules.
Owner:BOE TECHNOLOGY GROUP CO LTD

A method and system for evaluating skin color restoration perceptual color difference of a display device

The application discloses a kind of display equipment skin color restoration perceived color difference evaluation method and system, including obtaining reference image and to be evaluated image;Obtain preset reference white point coordinate;Reference image and to be evaluated image are converted from RGB color space to YCbCr color space, and skin color region pixel is extracted according to skin color threshold value;Skin color region RGB value is converted into CIE XYZ value and CIELAB value;The difference between the brightness of two images skin color region, chroma difference, hue difference, average brightness and average hue;Based on the fitting relationship of visual color difference and the average brightness, average hue of image pair of subjective visual evaluation experimental data, determine brightness direction weight parameter, chroma direction weight parameter and hue direction weight parameter, construct skin color difference evaluation model;Color difference parameter and weight coefficient are input into model, to obtain evaluation perceived color difference value, and output skin color restoration quality evaluation result.The application can be used for display equipment skin color restoration evaluation, white balance effect evaluation and related image processing scene.
Owner:UNIV OF SCI & TECH LIAONING