Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

270 results about "Associated image" patented technology

Fire situation analysis method based on multi-dimensional data fusion

The invention discloses a fire situation analysis method based on multi-dimensional data fusion, and particularly relates to the field of fire image analysis, and the method comprises the steps: analyzing the change of the direction retention rate between adjacent frames through extracting the main direction texture vectors of a building and a vegetation region, and recognizing object state change candidate segments; in the candidate area, combining main direction disturbance and image definition reduction to construct a spatial scoring graph, performing nonlinear amplification on the spatial scoring graph, and extracting a gradient increasing path to generate a structure damage main path set; then calculating a directional included angle between paths and a space coincidence rate, constructing a trend consistency aggregation channel graph, and extracting a spreading principal axis; and finally, superposing a temperature rise area in the thermal infrared image with a spreading principal axis, extracting a dual response area to generate a fire behavior boundary prediction layer, and realizing fire behavior spreading path prediction based on fire scene related image structure damage information analysis.
Owner:TIANJIN SHENGDA SECURITY TECH CO LTD +1

Object Detection Method and System Based on User-Defined Category

The provided is a method and system for object detection based on user-defined categories. The method includes: a user inputting a natural language description and a related image, obtaining a detection target auxiliary input using an auxiliary characterization generation technique for a detection target based on a phrase boundary point modeling technique; calling a detection target characterization generation model based on a multimodal reconstruction and alignment network to obtain a plurality of text characterizations of the detection target; generating target reverse characterizations based on an image-adaptive target characterization matching estimation technique to meet custom requirements of the detection target; and optimizing a vision-language multimodal model based on feedback data of the detection target of the user under detection, and optimizing the vision-language multimodal model based on the feedback data during usage of custom object detection.
Owner:HANGZHOU MEARI TECH CO LTD

Ai-driven creation of custom stickers from messages in chat interfaces

PendingUS20250378602A1Mathematical modelsNatural language analysisEngineeringVisual expression
This disclosure relates to techniques for generating and utilizing custom stickers in a digital communication environment. A technique involves receiving a text-based message input during a chat session and using a generative language model (e.g., a Large Language Model, or LLM) to create a text prompt. This prompt is then used by a generative image model to produce a custom sticker. The generated sticker is sent to a client device where it is displayed in a sticker tray alongside other selectable stickers. Users can select and send these stickers directly within their chat interface, enriching communication with visually expressive and contextually relevant imagery.
Owner:SNAP INC

Artificial intelligence chatbot

Methods and systems for interacting with users via a chatbot. A natural language query is received and processed by submitting a search query to a search engine. The search engine identifies relevant information including textual information and images for formulating a response. The identified information and query are submitted to a Large Language Model which generates a response displayed via the chatbot. The response may include textual information and relevant images. The system can extract text from images of documents and convert textual information into numerical vector representations for processing. Selectable options based on clustered relevant information can be provided to users for query refinement when appropriate. The chatbot interface enables natural language interactions while leveraging search capabilities and Artificial Intelligence to provide informative and helpful responses with both text and visual elements.
Owner:HONEYWELL INTERNATIONAL INC

Geosynchronization of an aerial image using localizing multiple features

A georegistration (a.k.a. georectification) of an image captured by a camera in an aerial vehicle, such as a satellite, is based on identifying multiple features using descriptor sets, and sending to a ground station only the descriptors of the identified features and the associated locations in the captured image, without sending of the captured image itself, thus requiring a low communication bandwidth. Using a database of geosynchronized reference images, the ground station uses the received descriptors sets and the associated image locations to localize the features on a selected geosynchronized reference image from the database, and forms a mapping function that map any locations in the captured image to geographical coordinates on Earth. The mapping may be used to geosynchronize an additional feature identified in the aerial vehicle, or to geo synchronize a region that may be cropped from the captured image and sent to the ground station.
Owner:EDGY BEES LTD

Explanatable farmland image enhancement method and system based on physical perception and reinforcement learning

ActiveCN120876343AImage enhancementImage analysisImaging processingIncrement threshold
The invention relates to the technical field of image processing, in particular to an interpretable farmland image enhancement method based on physical perception and reinforcement learning, and the method comprises the steps: obtaining an original farmland image, carrying out the global perception analysis, extracting multi-scale features, and recognizing a degradation region and feature distribution; on the basis of the global perception analysis result, constructing a semantic enhancement blueprint for quantifying the physical attributes and optimization requirements of the degradation area; initializing a reinforcement learning agent according to the semantic enhancement blueprint, and selecting an image processing operation sequence through a physical constraint reward function; executing the image operation sequence, and terminating enhancement processing according to the quality evaluation index increment threshold and the physical consistency of the semantic enhancement blueprint; and outputting an interpretability report of the enhanced image and associated image operation sequence physical basis traceability. The objective of the invention is to solve the technical problems of lack of interpretability, insufficient environmental adaptability and rigid decision-making mechanism of a farmland image enhancement technology.
Owner:CHINA TOWER CO LTD

Multi-modal sentiment analysis method and system based on thinking chain and background knowledge

The invention relates to a multi-modal sentiment analysis method and system based on a thinking chain and background knowledge. The method comprises the following steps: constructing a sample set containing original texts, related images, aspect items and sentiment polarities thereof; constructing a multi-modal sentiment analysis model, and generating background knowledge information by utilizing a visual language large model and combining an input text and an image; the generated background knowledge is screened, optimized and fused; for each sample in the sample set, in combination with the text data and the emotion polarity, generating a thinking chain thinking process and adding the thinking chain thinking process into the sample set, and then finely adjusting the visual language large model through the sample set added with the thinking chain thinking process; finally, the fused background knowledge, the text data and the picture data are sequentially input into the trained visual language large model, aspects in the text data are extracted, and the emotion polarity corresponding to the aspects is predicted. The method and the system are beneficial to improving the accuracy of multi-modal emotion polarity analysis.
Owner:FUZHOU UNIV

Tear total IGE detection method and system based on multi-modal data fusion

The invention discloses a tear total IGE detection method and system based on multi-modal data fusion. The method comprises the following steps: acquiring ocular surface images, physiological parameters and environmental allergen data, and acquiring historical medical data from a hospital information system; preprocessing and feature extraction are performed, patient symptom description is analyzed in combination with a natural language processing technology, and related image block features are recognized; adjusting a space-time attention weight based on frequency domain analysis and transfer learning, inputting a lightweight Transform model, and predicting a tear total IgE concentration and an allergy risk score; whether a trace tear sampling program is started or not is judged, and the actually measured concentration of total tear IgE is detected through a disposable micro-fluidic chip; and integrating all information by using a Bayesian adaptive filtering algorithm to determine the total tear IgE level. By implementing the method provided by the invention, the sampling amount is extremely small, the detection time is short, the environmental interference can be dynamically corrected, and the absolute concentration value is output, so that the dual requirements of clinical precise diagnosis and treatment and home monitoring are met.
Owner:SHANGHAI LIANGXIN TECHNOLOGY CO LTD

Extracting images and determining their meaning for semantic image retrieval and training a transformer-based multi-modal large language model to generate domain-aware images based on image meanings

The disclosure relates to systems and methods automatically extracting an image and related image components, computationally determining an understanding of the image, and generating mathematical vector embeddings via sentence encoders based on the computationally determined understanding. The mathematical vector embeddings may be used for semantic image retrieval that enables image searching based on a semantic understanding of input images and / or input text. The mathematical vector embeddings may be used for training and executing generative Artificial Intelligence (AI) models to create new content that includes retrieved images and / or generate new images.
Owner:ROHIRRIM INC

Deep learning technique for automated radiological image analysis and disease detection

A real-time artificial intelligence (AI) framework is provided for the automated analysis of radiological images and detection of disease, such as extracapsular extension (ECE) in prostate cancer. The system includes a dual deep learning architecture comprising a first convolutional neural network (CNN) for identifying diagnostically relevant image slices from three-dimensional MRI data, and a second CNN for classifying disease presence based on those slices. A preprocessing pipeline standardizes and harmonizes image input, and cropping algorithms isolate the region of interest for enhanced model performance. This framework enables scalable, high-accuracy diagnosis across various imaging modalities including but not limited to MRI, CT, PET, ultrasound, and diverse disease types, improving clinical decision-making and supporting integration into real-time radiology workflows.
Owner:RES FOUND THE CITY UNIV OF NEW YORK

Large-view-field high-resolution compound eye camera array system based on bionic human eye vision

The invention provides a large-view-field high-resolution compound eye camera array system based on bionic human eye vision. The system is composed of a curved surface support and an industrial camera array arranged on the curved surface support. Reverse extension lines of optical axes of all the cameras intersect at the same point to form a concave surface structure similar to the retina, and the fixed view field overlapping rate of the adjacent cameras is kept. The industrial camera array is used for collecting a target scene image, large-view-field information is obtained by fusing and splicing all camera images, and a high-resolution image is obtained by conducting super-resolution reconstruction on a plurality of related images in a view field center interested area. The system provides three working modes, namely a low-resolution mode, a conventional-resolution mode and a super-resolution mode, which can be switched by adjusting camera parameters. Finally, the system can perceive the dynamic change of the edge of the field of view in a high-frame-rate and low-resolution mode, magnify the details in the center of the field of view in a super-resolution mode, realize efficient non-uniform imaging, and optimize the information acquisition and processing efficiency and the structure compactness.
Owner:SICHUAN UNIV

Parking lot-based lidar monitoring method, system, device, and storage medium

The application provides a parking lot-based laser radar monitoring method, system, device and storage medium, wherein the method comprises the following steps: collecting point cloud data in a parking lot and dividing the point cloud data into a static point cloud set and a dynamic point cloud set, dividing target dynamic point cloud sets to be tracked from the static point cloud set and the dynamic point cloud set, obtaining a coincidence coefficient of a first projection area of a preset parking space and a second projection area of the target dynamic point cloud set, and updating the target dynamic point cloud set to a static point cloud when the target dynamic point cloud set is static in the range of a parking space point cloud and the coincidence coefficient meets a preset threshold, and determining that a collision is sent when the nearest distance between the second contour of the target dynamic point cloud set and the third contour of the static point cloud representing all vehicles parked on the parking space is 0. The application can solve the problem of car safety protection after parking, confirm the danger and capture the relevant image, and provide the greatest vehicle safety guarantee for the owner.
Owner:CHINA TELECOM CORP LTD

Monitoring video viewing method and device, computer equipment and medium

The invention relates to a monitoring video viewing method and device, computer equipment and a medium, and the method comprises the steps: determining a source camera of a target monitoring video needing to be viewed and corresponding viewing time information in response to a monitoring video viewing instruction; searching a camera to which an external monitoring video having a view overlapping area with the target monitoring video within the same time according to the viewing time information, and taking the camera as a part of neighbor cameras of the source camera; determining the optimal privacy metadata from the privacy metadata of the corresponding viewing time information of the partial neighbor cameras, wherein the privacy metadata comprises frame associated image blocks corresponding to the shielded areas of the privacy sensitive targets in the external monitoring videos of the corresponding neighbor cameras; and repairing a covered area of a corresponding privacy sensitive target in the target monitoring video according to the frame associated image block of the optimal privacy metadata, and then playing the target monitoring video. According to the invention, on the premise of default privacy protection, key details of the monitoring picture are intelligently restored and enhanced as required.
Owner:深圳市灵智无界科技有限公司

Applications for gain curves in imaging and video

Techniques are disclosed relating to exchange of images in networked computing applications. In particular, the disclosure relates to exchange of gain curves that are used to represent imaging and / or video in such applications. A gain curve may define a mathematical transformation that relates values from a source image domain to a destination image domain. The image and its associated gain curve(s) may be published to destination devices for consumption. When a destination device consumes the image, the destination device may apply a transform to source image content according to the gain curve(s) published with the image. For example, the destination device may apply a gain curve to an associated image directly, or it may derive another transform from the gain curve and additional information known to the destination device.
Owner:APPLE INC

Physician-guided machine learning system for assessing medical images to facilitate locating of a historical twin

A computer-implemented method of evaluating a user image of a patient to enable identification of a historical twin of the patient. The method includes organizing a plurality of medical images in an archive and receiving from a medical professional each of: (i) a region of interest; (ii) a textual description; (iii) selections for binary criteria; and (iv) weights of weighable criteria. The method comprises using a natural language search to create a relevant set of medical images and creating an optimal set from the relevant set of medical images by discarding medical images from the relevant set based at least on the selections for binary criteria. The method includes image processing medical images in the optimal set using the weight of the features of the region of interest to create medical image results. The relevant set comprises less than ten percent of the medical images in the archive.
Owner:IRANI NEVILLE

Region-text caption generation using global caption information

Approaches presented herein may be used to generate captions using raw caption information. Raw caption information may be used, with an associated image, to generate a detailed image caption. Object lists may then be generated from the image and / or the detailed image caption to produce an image including boxing box proposals for objects within the image. One or more trained machine learning systems may then be used to generate region of interest captions that infuse the global caption context associated with the raw caption information.
Owner:NVIDIA CORP

Aviation oil pipeline unmanned aerial vehicle intelligent inspection method and system

The invention discloses an aviation oil pipeline unmanned aerial vehicle intelligent inspection method and system, and the method comprises the steps: collecting visual image data along a pipeline through an inspection terminal carried by an unmanned aerial vehicle, covering a pipeline body and a surrounding environment, and recognizing an abnormal scene endangering the safety of the pipeline; inputting the image data into an abnormal scene recognition model deployed at an unmanned aerial vehicle end, and judging whether a preset type of abnormality is included; if at least one type of abnormity is identified, generating an alarm signal; alarm and related images are uploaded to a remote monitoring center server through wireless communication, and intelligent unmanned inspection of the running state of the pipeline is achieved. An image acquisition and recognition model is integrated at an unmanned aerial vehicle end, a pipeline and an environment are sensed in real time, and key abnormity is automatically recognized; the edge deployment model reduces invalid return, and only triggers alarm uploading when a risk is detected; and in combination with wireless communication return alarms, high-reliability monitoring is realized, and intelligent and refined guarantee is provided for safe operation of aviation oil pipelines.
Owner:CHINA AVIATION OIL PENGZHOU PIPELINE TRANSPORTATION CO LTD

Multi-modal map enhanced retrieval method and dialogue system based on feature fusion optimization

The invention discloses a feature fusion optimization-based multi-modal map enhancement retrieval method and a dialogue system. The method comprises the following steps of: respectively carrying out pre-training and fine tuning on a visual model and a language model by utilizing a domain image and text data; constructing a knowledge graph based on the text data in the knowledge base and constructing a vector database containing associated image data; performing semantic analysis and optimization on the original query of the user by using the language model and forming a structured retrieval intention; searching related sub-graphs, text semantic vector information and associated image data based on the search intention; encoding the sub-images into knowledge contexts, inputting the knowledge contexts into a dynamic prompt generator to generate visual prompts, and extracting enhanced visual features from the associated image data through a visual model; and inputting the subgraph, the text semantic vector information and the enhanced visual features into a language model for collaborative reasoning, and generating and outputting a final answer. According to the method, deep fusion and accurate retrieval of multi-modal knowledge can be realized, and the accuracy and efficiency are remarkably improved.
Owner:ZHEJIANG UNIV

Method, computer device, and computer-readable recording medium to provide message summary and associated image

A method of providing a message summary and an associated image may include requesting an image search in relation to a message summary created based on a message in a chatroom; receiving an image bundle that includes at least one image in response to an image search request; and displaying the image bundle in association with the message summary.
Owner:LINE PLUS

Power system-oriented dynamic knowledge base driven dialogue generation system, method, equipment and medium

The invention discloses a power system-oriented dynamic knowledge base driven dialogue generation system, method, equipment and medium, and the system comprises a dialect collection and recognition module which is used for collecting dialect voice data under different regional power scenes, carrying out the noise reduction of the dialect voice data, carrying out the dialect recognition through a deep learning model, and obtaining a dialect recognition result; converting the dialect voice data into a standard text; a dynamic knowledge base module; the language processing and image reasoning module is used for receiving the standard text, performing semantic understanding, generating semantic representation in combination with a knowledge base in the dynamic knowledge base module, and performing target detection and recognition on a power equipment fault related image input by a user to obtain an image recognition result; and a dialogue generation module. According to the invention, a cooperative system of four modules of dialect acquisition and identification, a dynamic knowledge base, language processing and image reasoning and dialogue generation is constructed, so that intelligent dialogue service oriented to the power industry is realized.
Owner:GUIZHOU POWER GRID CO LTD

Associated imaging impurity detection method and system for drug production

The invention discloses a correlated imaging impurity detection method and system for drug production, and relates to the technical field of drug impurity detection. The method comprises the following steps: loading a first binary mask matrix, projecting a target product area, controlling a single-pixel detector to detect, determining a first measurement value, and performing lightweight reconstruction to obtain a first preview; generating a second binary mask matrix and performing projection and detection reconstruction by pre-checking the first preview, and performing multi-round iteration until an Nth preview is determined; calling the first preview to the Nth preview from a temporary database of an online detection platform, executing multi-layer compressed sensing and reconstruction, and determining a reconstruction result; and aiming at a reconstruction result, matching in an impurity feature library, and determining an impurity verification result. The technical problems of low impurity detection efficiency and insufficient detection accuracy in the medicine production process in the prior art are solved, and the technical effect of efficient and accurate online detection of medicine impurities in the production link is achieved.
Owner:NANTONG MEDICAL DEVICES

Customization of vehicle-related images

A method for a driver assistant image customization of vehicle-related images, a data processing circuit, a computer program, a computer-readable medium, and a vehicle, can include, with at least one sensing device of the vehicle, obtaining an input image to be customized. With at least one human-machine-interface of the vehicle or connected thereto an input determining at least one customization scheme to be performed is received. Using at least one data processing circuit of the vehicle applying artificial intelligence the input image is transformed according to the at least one customization scheme into a transformed output image. With at least one smart mirror of the vehicle or a mobile device connected to the at least one data processing circuit the transformed output image is outputted. The at least one customization scheme includes a plurality of different types of adaptation modes.
Owner:FORD GLOBAL TECH LLC

system

A system is provided.SOLUTION: A system comprising: means for a user to upload an article; means for a server to parse the uploaded article; means for the server to automatically translate the parsed article into multiple languages; means for the server to extract and embed SEO keywords for the translated article; means for the server to generate and link relevant images and video to the article; and means for the server to distribute the processed article to platforms in each country.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

Image processing device, operation method of image processing device, and operation program of image processing device

An image processing device includes: a processor, in which the processor is configured to: obtain an evaluation value for a quality of an image; and determine a trimming method of a trimming target image, which is any one of the image or a related image of the image, based on the evaluation value.
Owner:FUJIFILM CORP

Driver assistance system and related image processing method

The invention relates to a driving assistance system comprising at least one optical sensor and at least one associated protection device (3) comprising a transparent and rotatably mounted optical element (5). The protection device (3) comprises at least one acquisition element (11) for acquiring at least one information describing at least one angular position of the optical element in the capture of an image. The system comprises a processing unit configured to detect the presence of at least one defect (100) in the optical element (5), to receive information describing the angular position of the optical element, to determine based thereon at least one angular position of said defect in the optical element and to remove said defect in the captured image by means of image processing, taking into account the determined angular position of said defect in the optical element.
Owner:VALEO SYST DESSUYAGE SAS

Poster background image selection, model training, poster generation method and related device

The application discloses a poster background image selection, model training, poster generation method and related device. Through the application of the technical solution, a plurality of weakly related image text pairs can be used to train a preset visual text model, and a visual text model obtained through the training is used to automatically select a poster background image weakly related to the text information of interest to a user. Then, a final poster image is generated based on the automatically selected poster background image. Thus, the problem that a large number of high-quality poster requirements cannot be met by relying only on manual design to generate posters in the related art is avoided.
Owner:BEIJING ACAD OF ARTIFICIAL INTELLLIGENCE +1

Three-dimensional reconstruction method and related apparatus

The embodiment of the application provides a three-dimensional reconstruction method and related devices. The method comprises: based on an image frame, extracting point data from point cloud data associated with the image frame and putting the point data into a target point data set; wherein, the associated image frame and the point cloud data are: there is an overlapping surface area between a surface area of a target object corresponding to the image frame and a surface area of the target object corresponding to the point cloud data; the point data put into the target point data set corresponds to the overlapping surface area; in the process of continuously scanning to generate the associated image frame and the point cloud data, point data meeting a specified condition in the point cloud data is added to the target point data set; and a three-dimensional reconstruction model of the target object is generated according to the point data in the target point data set. The three-dimensional reconstruction efficiency can be improved to a certain extent.
Owner:SCANTECH (HANGZHOU) CO LTD

Method for multispectral recording of an image stream and associated image recording system

A method for multispectral recording of an image stream, in which a sequence of single images of a scene, in particular a continuous video image data stream, is recorded as an image stream using an image sensor of an image recording system. At least two different types of single images are recorded in different associated wavelength ranges using the image sensor. At least one type A single image of the sequence is recorded during a chronological type A recording segment and at least one type B single image of the sequence is recorded during a chronological type B recording segment. The type A recording segment and the type B recording segment are chronologically separated from one another by a respective waiting interval, in which no image data are sensorially acquired using the image sensor. An image recording system is also provided that to carry out the method.
Owner:SCHOLLY FIBEROPTIC GMBH

Distributed camera system

A system having a central station and a plurality of cameras installed various locations. To search for and locate an item of interest, the central station generates and sends an item model to the cameras. When stored in a camera, the item model causes a logic circuit of the camera (e.g., a deep learning accelerator) to use image data, received from an image sensor for storing in a memory device of the camera, as an input to an artificial neural network. The logic circuit performs the matrix computation of the artificial neural network to generate a classification of whether the images are relevant to the item of interest characterized by the item model. If so, the camera transmits the relevant images to the central station for further processing to determine a real time location of the item of interest.
Owner:MICRON TECHNOLOGY INC

Audible auditing method and device, computer equipment, readable storage medium and program product

The invention relates to an added staff auditing method and device, computer equipment, a readable storage medium and a program product. The method comprises the following steps: when an auditing node is transferred to a first target node of a target workflow, acquiring an added personnel auditing associated image of a to-be-audited object, inputting the added personnel auditing associated image into an auditing model for extracting to-be-audited information to obtain auditing reference information, and finally, responding to an added personnel auditing operation for the to-be-audited object, and performing auditing on the to-be-audited object. And displaying the auditing reference information and the to-be-audited information of the to-be-audited object in a display interface, so as to determine an auditing result of the to-be-audited object at the first target node based on the auditing reference information and the to-be-audited information of the to-be-audited object. By adopting the method, the working efficiency of staff increase auditing can be remarkably improved.
Owner:CHINA LIFE INSURANCE CO LTD