Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

79 results about "Image representation" patented technology

Image-level representation is a (numerical) way to represent an image without a direct pixel representation. For example, one could represent an image by its histograms of luminance and chroma values, or by its Fourier transform, or by any other statistical measure. Such a representation helps compare images or detect specific features.

Vision positioning method based on memory self-correction

ActiveCN117668283BEnhancing Semantic Consistencyexact matchData setRadiology
The application discloses a visual positioning method based on memory self-correction. The existing visual positioning method uses fixed image and text representation to capture cross-modal semantic consistency, which limits the flexibility of adjusting image representation according to different text information. In order to cope with this limitation, the application proposes a new memory self-correction network, which dynamically refines the image representation according to the query, thereby improving the semantic consistency between the text and the image to realize visual positioning. A semantic related filtering module (SRFM) and an adaptive memory fusion module (AMFM) are constructed to explicitly model the relationship between the image and the text. SRFM focuses on filtering image information irrelevant to the query, while AMFM adaptively fuses text-related representation with initial image features to enhance the understanding ability of the MSCN model. Comprehensive experiments on three datasets verify the superiority of the proposed method compared with existing methods.
Owner:CHINA UNIV OF PETROLEUM (EAST CHINA)

A training and application method, device, and medium for an image reconstruction model.

PendingCN122312795AEngineeringComputational budget
This application relates to the field of image representation technology, and discloses a method, device, and medium for training and applying an image reconstruction model. The method involves selecting Gaussian sigmas during the image reconstruction model training process, combining the selected Gaussian sigmas into pairs, and fusing them into a new Gaussian sigma to replace the original two Gaussian sigmas. This proactive fusion mechanism effectively eliminates redundant representations in the model, freeing up storage and computational resources occupied by redundant Gaussian sigmas. This allows limited model capacity to be more concentrated on the complex regions in the image that truly require detailed modeling, thereby improving resource utilization and optimizing the overall image reconstruction quality.
Owner:SHENZHEN UNIV

Knowledge graph guided tin-based material multi-modal data organization method and system and readable storage medium

This application relates to the field of computer system technology, and more particularly to a knowledge graph-guided method for organizing multimodal data of tin-based materials, a computer system, and a readable storage medium. An initial knowledge graph of tin-based materials is constructed based on their composition, structure, and properties. Features are extracted from images of tin-based materials, and image quality is graded under the supervision of samples with quality level labels. An image representation and grading model is constructed, and the features of material nodes and their neighboring nodes in the tin-based material knowledge graph are learned. Finally, when the knowledge graph needs to be updated, the similarity between the features of newly added tin-based material images and the image features that may correspond to candidate material entities is compared to identify tin-based material entities with potential links. Triples for knowledge graph completion are constructed, achieving continuous improvement and effective organization of multimodal data of tin-based materials. The aim is to solve the problem of how to organize multimodal data of tin-based materials.
Owner:YUNNAN UNIV

Systems and methods for encoding and searching scenario information

Systems, methods, and non-transitory computer-readable media can receive a query specifying at least one example scenario. At least one image representation of the at least one example scenario can be encoded based on the query to produce at least one encoded representation. An embedding of the at least one representation of the at least one example scenario can be generated based on the at least one encoded representation. At least one scenario that is similar to the at least one example scenario can be identified based at least in part on the embedding of the at least one representation of the at least one example scenario and an embedding representing the at least one scenario. Information describing the at least one identified scenario can be provided in response to the query.
Owner:LYFT INC

A thyroid nodule ultrasound image feature extraction method and system

PendingCN122289718Aauxiliary judgmentimprove objectivityFeature extractionImaging analysis
This invention relates to the technical field of ultrasound image feature extraction for thyroid nodules, specifically to a method and system for ultrasound image feature extraction of thyroid nodules. The method includes the following steps: acquiring an ultrasound image containing thyroid nodules; performing multi-scale image analysis on the ultrasound image to construct a multi-scale image representation of the ultrasound image; using the multi-scale image representation to identify low-echo regions in the ultrasound image, and combining feature analysis, judgment correction, and spatial relationship verification to determine the acoustic shadowing region in the ultrasound image; performing suppression processing on the acoustic shadowing region, wherein the suppression processing includes unifying the signal within the acoustic shadowing region and softening the boundary of the acoustic shadowing region to obtain an optimized image; identifying the contour of the thyroid nodule based on the optimized image; and quantifying the morphology and internal structure of the thyroid nodule based on its contour. This application can effectively identify and suppress acoustic shadowing regions in ultrasound images.
Owner:THE FIRST AFFILIATED HOSPITAL OF MEDICAL COLLEGE OF XIAN JIAOTONG UNIV

A computing power network DDoS attack anomaly detection method and system based on traffic snapshot visual coding and time sequence behavior modeling

PendingCN122419910AAttackEngineering
The application discloses a kind of based on flow snapshot visual coding and timing behavior modeling computing power network DDoS attack anomaly detection method and system, first to original flow is carried out feature screening and normalization processing, and multidimensional flow feature is mapped into two-dimensional image representation, while retaining key space-time semantics Significantly reduce data dimension;Subsequently, a hybrid deep learning model is constructed, in which the flow snapshot visual coding feature extraction module based on EfficientNet efficiently captures the relevance between computing power task streams within a single time window, and its output is input into the LSTM module after structure remodeling as a time series to capture the flow timing evolution law caused by computing power scheduling;Finally, the spatiotemporal features are discriminated by a lightweight classifier after fusion.The present application fully meets the comprehensive needs of low resource overhead, high real-time and high accuracy of computing power network, and shows significant advantages in large-scale, high-bandwidth and distributed attack scenarios.
Owner:SUZHOU CHIEN SHIUNG INST OF TECH

Optical coherence tomography apparatus and computer program for displaying tomographic images

This invention provides an optical coherence tomography (OCT) apparatus and a computer program for displaying tomographic images. The OCT apparatus is a polarization-sensitive type. It includes an imaging unit and a display unit, wherein the imaging unit captures tomographic images of the examined eye; and the display unit displays the tomographic images captured by the imaging unit. The tomographic images include at least two of the following: an image representing the tissue within the examined eye using scattering intensity; an image representing the distribution of melanin within the examined eye; an image representing the fiber density within the examined eye; an image representing the direction of fiber travel within the examined eye; and an image representing blood flow within the examined eye. The display unit superimposes at least two images from the same location on the same cross-section. Therefore, the state of the examined eye can be readily and comprehensively assessed.
Owner:TOMY CO LTD

Systems and methods for predicting driving behaviors of drivers by transforming trip data into image representation

A computer-implemented method including receiving trip data of one or more trips of a driver from one or more sensors, dividing the trip data into a plurality of trip data segments based on a predetermined distance, transforming the plurality of trip data segments into a multi-dimensional graphical representation beyond two dimensions by generating a two-dimensional graphical representation using relative longitude and latitude coordinates extracted from the trip data as indexes and adding depth to each point of the two-dimensional graphical representation to form a high-depth image-like tensor, and determining predicted driving behaviors of the driver by inputting the multi-dimensional graphical representation into a prediction model to extract features indicative of driving behaviors. Other embodiments are described.
Owner:QUANATA LLC

Image generation method and device based on spatiotemporal data interaction and electronic equipment

The application provides an image generation method and device based on space-time data interaction and electronic equipment, relates to the technical field of computer vision, and aims to realize efficient image generation. The method comprises the following steps: acquiring an image representation sequence of the last time; the image representation sequence comprises a visible representation and a mask representation, the mask representation represents unknown image content, and the visible representation represents known image content, which is used for providing image information for the mask representation to infer unknown image content; performing self-attention interaction-based coding on the visible representation to obtain a visible representation feature; performing cross-self-attention interaction-based decoding on the mask representation according to the visible representation feature to obtain a new image representation sequence; performing image representation sequence iteration generation for multiple times according to the above steps; and generating an image according to the new image representation sequence in the case that the new image representation sequence does not contain the mask representation.
Owner:TSINGHUA UNIVERSITY

X-ray CT apparatus, method, and program

Even if the object is curved in its longitudinal direction, its internal structure can be analyzed with high precision. [Solution] The X-ray CT apparatus of this disclosure comprises a first generation unit 110, a point cloud determination unit 115, and a second generation unit 120. The first generation unit 110 generates a three-dimensional CT image of the object. The point cloud determination unit 115 determines first to M points (M is a natural number greater than or equal to 2) that represent the central axis of the object. The second generation unit 120 generates CTS images 300_1 to 300_M according to the three-dimensional CT image. The CTS image 300_i (1≦i≦M) is an image representing the cross-sectional region of the object in the i-th plane, which is orthogonal to the i-th vector that points from the i-th point to the j-th point (i≠j, 1≦j≦M) in its neighborhood. The k-th point (2≦k≦M) is located a predetermined distance from the k-1 point in the direction of the k-1 vector.
Owner:SUMITOMO ELECTRIC INDUSTRIES LTD

Gaming machine

To provide gaming machines that make the game easy to understand. [Solution] A gaming machine comprising: a win / fail lottery means that performs a win / fail lottery triggered by the fulfillment of predetermined conditions; and an effect execution means that can perform a variation effect from the time the decorative pattern displayed in the display area starts to vary until it stops in a manner corresponding to the win / fail lottery result by the win / fail lottery means, wherein a normal background image and a special background image that has a higher probability of indicating a win / fail result than when the normal background image is displayed are provided as background images displayed as the background of the decorative pattern in the variation effect, a special background image X is provided as the special background image, and when the variation effect in which the special background image X is displayed is a loss, and the normal background image is displayed in the next variation effect, a specific transition image can be displayed.
Owner:SANSEI R&D KK

Ophthalmic observation apparatus, method of controlling the same, and recording medium

PendingUS12672776B2OphthalmologyDisplay device
An ophthalmic observation apparatus according to an embodiment example includes a moving image generating unit, an analyzing processor, and a display controller. The moving image generating unit is configured to generate a moving image by photographing a subject's eye into which an artificial object has been inserted. The analyzing processor is configured to analyze a still image included in the moving image to identify a first site image corresponding to a predetermined site of the subject's eye and a second site image corresponding to a predetermined site of the artificial object. The display controller is configured to display, on a display device, the moving image, first position information that represents a position of the first site image, and second position information that represents a position of the second site image.
Owner:TOPCON CORPORATION

Gaming machine

PendingJP2026110741AEngineeringGame machine
This gaming machine helps prevent players from mistakenly identifying an unattended machine as an empty one, thereby reducing the occurrence of problems. [Solution] The image display means of the gaming machine can display an "away" image indicating that the player is away from their seat in response to a predetermined operation by the player in a predetermined situation. The predetermined situation includes a situation in which no fluctuation display is being performed by the fluctuation display means and the player can bet the game value they possess by betting, and a situation in which the bet has already been placed. Furthermore, it is possible to terminate the display of the "away" image in response to a specific operation performed on the performance operation means. In addition, while the "away" image is being displayed and the player is in a bet state, the display of the "away" image can be terminated when the start operation detection means detects a start operation, and the fluctuation display by the fluctuation display means is started in response to the detection of the start operation.
Owner:UNIVERSAL ENTERTAINMENT CORP

Non-transitory computer-readable storage medium storing program, point selection method, and information processing apparatus

A program includes causing a processing apparatus to acquire a first captured image representing a result of imaging a range including a first display area that has a corner, the first display area being located on a projection surface on which a first projection image is displayed by a first projection apparatus, detect, by performing image processing on the first captured image, a plurality of first feature points corresponding to a corner of a first image indicating the first display area in the first captured image, and display, by controlling a display apparatus, a first superimposed image obtained by superimposing at least a part of a plurality of first display images selected by a user on the first captured image, the plurality of first display images corresponding one-to-one to the plurality of first feature points.
Owner:SEIKO EPSON CORP

Image manipulation with sparse control

A method, apparatus, non-transitory computer readable medium, and system for image processing include obtaining an input image and a modification input, wherein the input image depicts an object and the modification input indicates a change to the object. A feature map is generated, and the feature map represents the object based on the input image. The feature map is transformed to obtain a transformed feature map based on the modification input. The transformed feature map represents the change to the object. A synthetic image is generated, using an image generation model, based on the input image and the transformed feature map. The synthetic image depicts the change to the object.
Owner:ADOBE INC

VISUALIZATION OF SEALED VESSELS

UndeterminedDE112024003977T5Motion vectorRadiology
A system for providing an image representation of an occluded vessel is provided. The system comprises one or more processors configured to: detect a segment (140p, 140d) of the vessel adjacent to the occlusion (130) in a sequence of angiographic images (120); extract, from the sequence of angiographic images (120), motion vectors (150) representing movement of the segment (140p, 140d) of the vessel based on the detected segment; mutually register the sequence of angiographic images (120) based on the extracted motion vectors (150); process the image intensities in the mutually registered sequence of angiographic images in an area of ​​interest adjacent to the segment (140p, 140d) of the vessel to provide the image representation of the occluded vessel; and output the image representation of the occluded vessel.
Owner:KONINKLIJKE PHILIPS NV

Method and apparatus for autonomously plugging a charging plug into a charging socket of a vehicle

Technologies and techniques for autonomously plugging a charging plug into a charging socket of a vehicle, wherein the charging plug is installed on a programmable robot arm of a charging station. A camera is used to locate the charging socket using image analysis of acquired images of the charging socket and the surroundings of the charging station are represented as a digital map including a charging socket region in which the charging socket is located. The charging plug is moved into a plug-in position at a distance from the charging socket and the digital map is processed into a modified digital map having a remote charging socket region. A plugging-in operation is then performed based on the modified digital map.
Owner:VOLKSWAGEN AG

A medical image representation learning method and system based on multi-granularity world modeling

The application provides a medical image representation learning method and system based on multi-granularity world modeling. The method first enhances the radiograph image into first and second enhanced views with spatial overlap, and inputs a visual transformer to extract basic patch features; then constructs multi-granularity anatomical representation through hierarchical aggregation; then drives the world model to perform anatomical structure modeling, anatomical layout modeling and domain change perception modeling tasks, infers the relative spatial coordinates across views by using the overlap ratio of the overlapping area features, and modulates the input features by using the granularity perception enhancement parameters; finally, the model parameters are optimized based on a joint loss function. The application can solve the technical problems of the prior art, such as lack of unified modeling of multi-level anatomical semantics of the radiograph image, insufficient cross-view spatial layout reasoning capability, and difficulty in maintaining anatomical consistency under domain change.
Owner:CHINESE PEOPLES LIBERATION ARMY ARMY SPECIAL MEDICAL CENTER +1

A train positioning method and positioning device

The application discloses a train positioning method and a positioning device. The positioning device acquires a plurality of code images obtained by photographing a plurality of position marks arranged on a track during train driving; for each code image in the plurality of code images, the following operation is performed to determine the position of the train represented by each code image: identifying the first code image to determine a first code result of the first code image; determining a first position of the first code image according to a correspondence relationship between the first code result and a pre-set code result and position; wherein the first code image is any one of the plurality of code images; and positioning the train according to a plurality of positions corresponding to the plurality of code images. The positioning accuracy is improved while the positioning cost is reduced.
Owner:青岛佳都微联信号系统有限公司

Robot action control method, device and electronic equipment

The application provides a robot action control method and device and electronic equipment, the action control method comprises: encoding natural language operation instructions and image data to obtain image feature sequence, target embedding vector and text embedding vector; performing dot product operation based on embedding retrieval on the image feature sequence and the target embedding vector to obtain a similarity score matrix, and performing mask reconstruction on the image feature sequence to obtain a fine-grained perception feature sequence; performing feature vector linear modulation aggregation on the image feature sequence and the text embedding vector, and performing pruning routing on the fine-grained perception feature sequence to obtain a target visual feature sequence; based on the target visual feature sequence, outputting an action trajectory sequence, and controlling a target robot to perform a corresponding action by an action expert layer. Through the above method, the accuracy of image representation and the processing capability of complex graphics are improved, and the logical close coupling of the action trajectory sequence and the operation instructions is ensured.
Owner:58 INTELLIGENT TECH (HANGZHOU) CO LTD

Gaming machine

PendingJP2026110740AEngineeringGame machine
This gaming machine helps prevent players from mistakenly identifying an unattended machine as an empty one, thereby reducing the occurrence of problems. [Solution] The image display means of the gaming machine can display an "away" image indicating that the player is away from their seat in response to a predetermined operation by the player under predetermined conditions. Furthermore, the administrator can select whether to enable or disable the display function of the "away" image in the administrator menu. The predetermined conditions include the state where a bet has been placed. In addition, the display of the "away" image can be terminated in response to a specific operation performed on the performance operation means. Moreover, while the "away" image is being displayed and the player is in the state where a bet has been placed, the display of the "away" image can be terminated when the start operation detection means detects a start operation, and the variation display means starts displaying a variation in response to the detection of the start operation.
Owner:UNIVERSAL ENTERTAINMENT CORP

Computational lithography simulation using a multi-channel physics-informed neural network for a 3D mask

PCT designated stageWO2026139198A1Lithography processAlgorithm
A non-transitory computer-readable medium stores a set of instructions that is executable by at least one processor of an apparatus to cause the apparatus to perform operations for simulating a lithography process. The operations include obtaining a physics-informed neural network (PINN) comprising multiple channel. A first channel of the multiple channels is associated with a first set of physical characteristics of an interaction between an electromagnetic (EM) field and a thick mask used in the lithography process. A second channel of the multiple channels is associated with a second set of physical characteristics of the interaction that is different from the first set of physical characteristics. The operations also include executing the PINN using the first channel to generate a near-field image representation of the EM field for the thick mask with the second channel disabled.
Owner:ASML NETHERLANDS BV

Systems and methods for dynamic server control based on estimated script complexity

A computer system includes processor hardware and memory hardware storing instructions for execution by the processor hardware. The instructions include, in response to receiving a first script from a user device, compiling the first script, generating an image representation of the compiled first script, and determining an estimated runtime of the first script using a machine learning algorithm. The instructions include transmitting the estimated runtime for display on a display of the user device, categorizing the estimated runtime, and transmitting the first script to a queue based on the categorization. The instructions include, in response to the first script reaching a front of the queue, executing the first script on a server of the plurality of servers that corresponds to the queue. The instructions include, in response to the first script being executed, transforming the display of the user device according to instructions of the first script.
Owner:CHARLES SCHWAB & CO INC

A vulnerability detection method based on a hierarchical centrality fusion strategy and a double-channel convolutional neural network

The application provides a vulnerability detection method based on a hierarchical centrality fusion strategy and a double-channel convolutional neural network, and belongs to the technical field of source code vulnerability detection. The method comprises the following steps: 1, denoising and normalizing the source code function for preprocessing, to generate a standardized code text; 2, parsing the standardized code text into a program dependency graph, and weighting and integrating the code semantic features and the centrality structure features of the graph nodes; 3, based on the hierarchical centrality fusion strategy, the integrated feature matrix is hierarchically mapped according to the local, propagation and global topological properties, to generate a multi-view RGB image representation; and 4, a double-channel convolutional neural network is constructed, deep features of the multi-view RGB image are extracted in parallel, and the vulnerability classification detection is completed based on the fused enhanced features.
Owner:JIANGSU UNIV

A method and system for controlling the brightness of an LCD display

The present application relates to display control technical field, specifically to a kind of LCD display brightness control method and system, method is applied to the display device containing direct backlight module and liquid crystal panel, backlight module is the multi-partition light source of independently adjustable brightness, method includes: constructing backlight module diffusion simulation module, based on point spread function to backlight brightness distribution diffusion obtains backlight output image;Utilize training sample to train HDR reconstruction model and local dimming model, HDR reconstruction model is based on saturation discrimination generation feature mask and is propagated in convolution feature extraction to inhibit saturated area feature, local dimming model generates supervision signal by means of diffusion simulation module and is updated with the loss function containing image fidelity term and power consumption constraint term;Display stage carries out luminance correction to input low dynamic range image and combines feature mask and exports target HDR image;Again output backlight prediction image and obtain backlight output image by diffusion simulation;According to the partition topological partition convergence forms backlight control image, and by target HDR image and backlight output image generates liquid crystal panel control image representation pixel transmittance;Accordingly drive each partition luminous intensity and each pixel transmittance, realize backlight and liquid crystal collaborative modulation output display image.The present application generates backlight partition control and liquid crystal panel transmittance control, so that display output approximates target HDR image and gives consideration to power consumption constraint and halo artifact suppression.
Owner:SHENZHEN FWS TECH CO LTD

Image processing model acquisition, image processing method and device

The application relates to the technical field of image processing, and discloses an image processing model acquisition method and device and an image processing method. In the method, for each sample image in a sample image set, a visual feature extraction network is used to extract the visual feature of the sample image; a text feature extraction network is used to extract the text feature of the sample image; a feature processing network is used to determine a first fusion feature and a first semantic matching score based on the visual feature and the text feature; a scoring network is used to determine an image representation score based on the first fusion feature and the first semantic matching score; and a first model is subjected to parameter adjustment based on a first loss function and a second loss function until a target image processing model is obtained. The above method is beneficial to improving the representation capability and accuracy of the image representation score output by the model.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Generating deep-linked stochastic images

Methods and systems are described herein for generating deep-linked stochastic image representations of access tokens that embed token access deep links on a mobile application interface. The system may obtain, in connection with a request to register an access token with an account, token data associated with the access token and event data associated with one or more events performed with the access token. The system may generate, for input to a stochastic machine learning model, input vectors using the token data and the event data. The system may obtain, via the stochastic machine learning model based on the input vectors, an image for the access token and may generate, for display on a user interface associated with the account, an image representation of the access token including the image and a deep link to functionality associated with the access token.
Owner:CAPITAL ONE SERVICES LLC