Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

2159 results about "Digital image" patented technology

A digital image is a numeric representation, normally binary, of a two-dimensional image. Depending on whether the image resolution is fixed, it may be of vector or raster type. By itself, the term "digital image" usually refers to raster images or bitmapped images (as opposed to vector images).

Three-dimensional reconstructions based on gaussian primitives

In implementation of techniques for three-dimensional reconstructions based on Gaussian primitives, a computing device implements a reconstruction system to receive a first digital image depicting an object from a first angle and a second digital image depicting the object from a second angle. The reconstruction system segments the first digital image and the second digital image into patches. The reconstruction system then generates, using a machine learning model, three-dimensional Gaussian primitives that predict parameters of points of the object in a three-dimensional space that correspond on a per-pixel basis to pixels of the patches. The reconstruction system then forms a three-dimensional reconstruction of the object for display in a user interface by merging the three-dimensional Gaussian primitives.
Owner:ADOBE INC

Image processing method applied to printed matter surface color difference detection

The invention discloses an image processing method applied to printed matter surface color difference detection, which relates to the technical field of image processing, and comprises the following steps: acquiring a digital image of a printed matter to be detected, performing illumination non-uniformity correction, and converting the corrected digital image into a CIELAB color space; performing region segmentation on the digital image based on double constraint conditions of color gradient and texture boundary to generate a detection region graph; feature parameters are extracted based on the detection area graph, and a feature data set is generated; establishing a dynamic reference model by utilizing process parameters and material characteristics of the printed matter; and calculating the distance between the feature data set and the dynamic reference model, identifying color difference regions and generating color difference scores, and screening and grading the color difference regions according to the color difference scores. According to the method, a multi-layer image pyramid and homomorphic filtering combined illumination correction technology is adopted, and an adaptive weight fusion mechanism based on image features is introduced, so that the illumination nonuniformity is effectively eliminated while the definition of printing details is kept.
Owner:GUANG ZHOU BEIDE PACKAGING & PRINTING CO LTD

Wound surface identification charging system based on precision medical treatment

The invention belongs to the technical field of medical image processing, and discloses a wound identification and charging system based on precision medical treatment, which realizes intelligent management of wound treatment through cooperative work of multiple modules. The method comprises the following steps: firstly, collecting a multi-angle digital image of a wound surface, performing standardization processing, performing image enhancement and noise reduction, and then performing identification and segmentation by using a deep learning algorithm; and performing three-dimensional reconstruction based on the accurate contour map to form parameter feature vectors, and performing intelligent wound type discrimination and severity evaluation. And according to an evaluation result, matching an optimal dressing change type and predicting required consumables, generating a precise dressing change scheme, and further calculating personalized charges and performing transparency verification. And finally, the system automatically generates a standardized medical service file through historical data comparison and rationality analysis. According to the system, the precision, standardization and transparency of wound treatment are realized, the medical service quality is improved, and the resource allocation is optimized.
Owner:THE SECOND AFFILIATED HOSPITAL OF GUANGZHOU MEDICAL UNIVERSITY

Processing multi-type document for machine learning comprehension

A computer-implemented method includes receiving a digital image of a document and a workflow describing an automation task. The method also include converting the digital image of the document into rich text that includes layout information in the document. The method further includes creating, based on the rich text and the workflow, a tree of thoughts that includes nodes and edges connecting the nodes and that binds at least some nodes representing the rich text with a task node representing the automation task. The method also includes converting the nodes and edges of the tree of thoughts into a natural language text. The method further includes inputting the natural language text into a language machine learning model with attention given to a token in the natural language text representing the task node. The language machine learning model, in response, outputs a result of completing the automation task.
Owner:INTERNATIONAL BUSINESS MACHINE CORPORATION +1

Building vision-language models using masked distillation from foundation models

The present disclosure relates to systems, non-transitory computer-readable media, and methods for training and implementing a vision-language model using masked distillation and contrastive image-text training. In particular, in one or more embodiments, the disclosed systems generate, utilizing a vision encoder, an image embedding from a masked digital image comprising a digital image with one or more masked patches. In some embodiments, the disclosed systems generate, utilizing a text encoder, a text embedding from a masked text phrase. In one or more embodiments, the disclosed systems generate, utilizing the vision-language model from the image embedding and the text embedding, a predicted text reconstruction of the text description and a predicted image reconstruction of the digital image. In some embodiments, the disclosed systems modify parameters of the vision-language model according to a masked distillation loss between the predicted text reconstruction and a text reconstruction generated by a pretrained large language model.
Owner:ADOBE INC

Deformation online measurement and control method for multi-robot collaborative assembly

The invention relates to a deformation online measurement and control method for multi-robot collaborative assembly. The method comprises the steps that multi-view surface images of workpieces in the assembly process are collected in real time through distributed robots and cameras arranged at the tail ends of the distributed robots; each view angle surface image is input into a deep learning model trained based on a digital image related technology, the optimal pixel displacement corresponding to each view angle surface image is output, and the optimal pixel displacement corresponding to each view angle surface image is converted into a spatial displacement label; the deep learning model takes an encoder-decoder as a trunk network; fusing the spatial displacement labels corresponding to the view angle surface images to obtain a fused displacement field; based on the fusion displacement field and the nominal path planning point, the tail end pose of the distributed robot is determined; the assembly error is calculated based on the reference target positioning and the tail end pose, and the PID controller adjusts the joint space of the distributed robot based on the assembly error. According to the method, the calculation overhead is remarkably reduced, and the measurement precision and robustness are improved.
Owner:HUNAN UNIV

Track generation method for unmanned aerial vehicle to track and aerially photograph target vehicle

The invention relates to the field of digital image detection and signal processing, and particularly discloses an unmanned aerial vehicle tracking aerial target vehicle trajectory generation method, which comprises the following steps of: constructing a moving target detection neural network model, and performing stage processing on micro, medium and fast moving optical flow features on an input image by the model through a hierarchical cascade optical flow attention mechanism to obtain a moving target detection neural network model; motion processing of video frames is improved using bidirectional timing optical flow enhancement. Inputting a target vehicle video into a model to obtain a center coordinate of a target vehicle detection frame in each frame as a position coordinate of a vehicle, and connecting the position coordinates according to a time sequence to form a preliminary track; performing decoupling compensation of the motion of the unmanned aerial vehicle on the initial track through a multi-scale adaptive dense optical flow algorithm; performing coordinate transformation to obtain a roughly estimated trajectory of the trajectory after decoupling compensation in a geodetic coordinate system; and identifying an abnormal frequency through wavelet transform, removing noise points by using Lagrange interpolation, and de-noising by applying extended Kalman filtering to generate an accurate trajectory of the target vehicle.
Owner:NANJING UNIV OF AERONAUTICS & ASTRONAUTICS

Image tampering detection method and system based on mixed features and RGB features

The invention relates to the technical field of digital image security and authentic identification, and provides an image tampering detection method and system based on mixed features and RGB features, and the method comprises the steps: obtaining a to-be-detected input image, and carrying out the preprocessing of the to-be-detected input image; respectively extracting a Haar wavelet high-frequency component, a discrete cosine transform frequency domain feature and a Bayer convolution noise feature, and carrying out matrix level fusion to obtain a mixed feature; extracting RGB (Red, Green and Blue) features for the preprocessed input image; the mixed features are connected through cross-layer residual errors, and mixed feature learning features are obtained; and integrating the mixed feature learning features and the fused RGB features by using a cross-modal feature interaction architecture to obtain a prediction probability graph. Multi-modal features are fused, high-frequency response is enhanced, and the accuracy of image tampering detection is improved by adopting a dynamic fusion mechanism. The technical problems that an existing tampering detection method is insufficient in feature characterization capacity in a complex scene, low in tampering trace detection sensitivity and the like are solved.
Owner:SHANDONG UNIV

Intelligent agent digital image interaction generation method based on multi-modal perception

The invention discloses an intelligent agent digital image interaction generation method based on multi-modal perception, which comprises the following steps: collecting multi-modal input data of a user, and respectively carrying out preprocessing and feature extraction on the multi-modal input data; inputting to an improved efficient modal cross learning network, and carrying out multi-modal feature fusion processing; constructing a semantic intention map, introducing a time index edge weight and an emotion driving edge weight, and encoding the map by using a structure perception map neural network; a modal style vector is extracted through a cross-modal style contrast learning mechanism, and a personalized style coding vector is generated through a hierarchical nested structure; inputting a personalized regulation and control gating mechanism, and regulating and controlling the middle layer representation in the interaction strategy generation process by adopting a feature channel linear modulation method; inputting the representation vector into a behavior strategy generation module to generate a multi-modal behavior output sequence; and the sequence is output to drive the digital image to perform synchronous response, and natural response generation in the user interaction process is completed.
Owner:JIANGSU ELECTRIC POWER INFORMATION TECH

Grain loss detectors for a combine harvester

A grain loss detector disposed on a combine harvester to detect grain loss within crop material being discharged by a separating or cleaning stage component of combine harvester while harvesting a crop. The grain loss detector includes at sensor generating sensor data output. A computing device receives the sensor data output for processing via software which compares the sensor data output against crop data characteristics to identify any grain kernels from among material other than grain. In one embodiment the sensor is a camera and the generated sensor data output is a digital image frame capturing the grain kernels and material other than grain passing the camera.
Owner:BUSHEL PLUS LTD

Generating digital images utilizing a diffusion-based network conditioned on lighting-aware feature representations

Methods, systems, and non-transitory computer readable storage media are disclosed for generating digital images with a diffusion-based generative neural network conditioned on background-extracted lighting features. The disclosed system determines, in response to a request to generate a digital image, a target background image for inserting a foreground object into the target background image. The disclosed system generates, from the target background image and utilizing a lighting conditioning neural network, a lighting feature representation indicating one or more lighting parameters of the target background image. Additionally, the disclosed system generates, utilizing a diffusion-based generative neural network conditioned on the lighting feature representation, the digital image including the foreground object inserted into the target background image based on a composite image comprising the foreground object and the target background image with a foreground mask corresponding to the foreground object.
Owner:ADOBE INC

Torque sampling inspection method and sampling inspection system

The invention relates to the technical field of torque sampling inspection, in particular to a torque sampling inspection method and system. Comprising the following steps: tool identity digitization: performing laser etching on an encrypted two-dimensional code in a tool stress concentration area, and associating and storing tool specification parameters, detection standards and use constraint conditions to a central database; intelligent task allocation: the system generates a dynamic detection plan according to a production line real-time state, an equipment service cycle and an operator skill level; multi-source data acquisition: acquiring tool identity information through a scanning device, and synchronously acquiring a physical signal of a torque detection device and a digital image of a tool identification file; grading judgment processing: performing process conformity verification on the measurement data, and generating a process stability evaluation report in combination with historical detection records; according to the closed-loop control strategy, the detection scheme is automatically adjusted according to the quality fluctuation characteristics. According to the scheme, the problem that a traditional method is seriously insufficient in the aspects of detection efficiency, abnormal traceability, process control and the like can be solved.
Owner:GUANGZHOU KERIS INFORMATION TECHNOLOGY CO LTD

Latent space based steganographic image generation

Techniques for latent space based steganographic image generation are described. A processing device, for instance, receives a digital image and a secret that includes a bit string. A pretrained encoder of an autoencoder generates an embedding of the digital image that includes latent code. A secret encoder is trained and utilized to generate an embedding of the secret to act as a latent offset to the latent code. The processing device leverages a pretrained decoder of the autoencoder to generate a steganographic image based on the embedding of the secret and the embedding of the digital image. The steganographic image includes the secret and is visually indiscernible from the digital image. Further, the processing device is configured to recover the secret from the steganographic image, such as by training and leveraging a secret decoder to extract the secret.
Owner:ADOBE INC

Fabricated building construction whole process management system based on artificial intelligence

The invention relates to the technical field of building construction management, and discloses an assembly type building construction whole process management system based on artificial intelligence. The system comprises a construction element sensing layer, a construction state mapping layer, a strategy synthesis layer and an execution coordination layer which are connected in sequence. The construction element sensing layer continuously captures on-site multi-modal environment data and component state data; the construction state mapping layer fuses multi-source data into a construction scene digital image with space-time relevance, and key path nodes and potential conflict domains in a construction process are analyzed and identified through a depth mode; the strategy synthesis layer generates a dynamic optimization strategy set for hoisting, transportation and installation processes based on the identification result; and the execution coordination layer drives on-site construction equipment to execute the corresponding instruction sequence. By constructing the digital image and deeply analyzing the construction mode, accurate insight and dynamic optimization of the construction process are realized, the construction efficiency and the automation level are improved, and process conflicts and resource waste are reduced.
Owner:FUJIAN POLYTECHNIC OF WATER CONSERVANCY & ELECTRIC POWER

Detecting problems of an auxiliary light beam of a laser surgical system

In certain embodiments, ophthalmic laser system includes an auxiliary light system, laser system, imaging system, and computer. The auxiliary light system directs auxiliary light towards a test target according to a planned test pattern to yield one or more actual auxiliary spots of an actual test pattern on the test target. The planned test pattern indicates one or more planned auxiliary spots located relative to a planned laser spot in a predetermined manner. The laser system directs a laser beam to yield an actual laser spot of the actual test pattern on the test target. The imaging system generates a digital image of the actual test pattern. The computer analyzes the digital image to compare the actual to the planned test pattern, detect a deviation of the actual from the planned test pattern, identify an issue indicated by the deviation, and provide output in response to the issue.
Owner:ALCON INC

Thermal shock failure prediction and laser repair device for thermal barrier coating and application method

The invention discloses a thermal shock failure prediction and laser repair device for a thermal barrier coating and an application method. The device comprises a multi-physics field coupling simulation analysis module, a thermal shock experiment loading module, an online damage monitoring and failure diagnosis module, a robot laser repair execution module and a central control and data processing core unit. A thermal-mechanical-chemical coupling model is established to predict a coating failure behavior, a high-frequency induction heating and gas quenching system is adopted to simulate an extreme thermal shock environment, a multi-band thermal infrared imager, an ultrasonic probe and a digital image related system are integrated to monitor damage evolution in real time, and in-situ repair of a damaged area is realized based on a six-degree-of-freedom industrial robot. According to the method, the whole-process closed-loop control from failure prediction to repair regeneration is realized, the problems of single function and lack of multi-field coupling analysis and real-time repair capability in the prior art are solved, and the service reliability evaluation precision and the remanufacturing efficiency of the thermal barrier coating are remarkably improved.
Owner:WUHU INST OF TECH

Method for reconstructing bedding structure rock finite element model based on rock core digital image

The invention provides a bedding structure rock finite element model reconstruction method based on a rock core digital image, and belongs to the technical field of digital rock cores, and the method comprises the following steps: S1, obtaining a gray level image of a rock sample; s2, calculating a transverse mean value of the analysis area to obtain a longitudinal gray level distribution curve; s3, applying moving average filtering and smooth filtering to obtain a smooth gray curve; s4, first-order difference is carried out on the smooth gray level curve, the position with the absolute difference value larger than a threshold value is taken as an initial stripe boundary candidate, and starting and stopping pixels, the width and the average gray level of each stripe are determined; s5, mapping each clustering label interval into a physical interval based on FEM grid partition; and S6, constructing a sample finite element model, and generating a reconstructed finite element model with real partition information. According to the method, the macrostructure features of the rock can be efficiently recognized and extracted by directly utilizing the pictures and combining gray profile analysis, and efficient and low-cost reconstruction of the finite element model of the bedding structure rock is achieved.
Owner:CHINA UNIV OF PETROLEUM (EAST CHINA)

Apparatuses and methods for training and using computational operations for digital image processing

An apparatus and method for training and using a computing operation for digital image processing are provided. The apparatus and method may be used for 3-dimensional medical images. An exemplary method for digital image processing comprises: receiving an image displaying at least one detectable structure, determining the detectable structure; segmenting the image to obtain a segmentation mask that is associated with a geometric shape and comprises at least one quantifiable visual feature; generating a mesh based on the quantifiable visual feature; computing at least on quantifiable visual parameter based on the mesh; extracting quantifiable visual data from the image based on the quantifiable visual parameter; training the computing operation with the quantifiable visual data. The method for digital image processing further comprises: receiving another image; segmenting, generating a mesh, computing quantifiable visual parameters, and extracting quantifiable visual data; and classifying the extracted quantifiable visual data with the trained computing operation.
Owner:MEDIAN TECH

Stay cable force rapid measurement method and system based on digital image correlation method

The invention discloses a method and a system for rapidly measuring cable force of a stay cable based on a digital image correlation method. The method comprises the following steps: acquiring video data of the stay cable; a line segment measurement calibration method is adopted, stay cable texture points serve as feature points, and a feature point selection mechanism is established; establishing a pixel-physical space coordinate mapping model based on a redundant observation method; model parameters are calculated through a collaborative calibration method; performing displacement measurement based on a digital image correlation method; establishing a double-layer feature extraction mechanism to complete local and global feature extraction; weight distribution of the local features and the global features is dynamically adjusted through a weight adaptive fusion strategy; and finally, establishing a double-layer tracking and failure recovery mechanism to complete hierarchical collaboration of coarse-grained global tracking and fine-grained local tracking. And carrying out fast Fourier transform on the dynamic displacement measurement result, and calculating to obtain the cable force. And reliable technical support is provided for non-contact accurate calculation and measurement of the cable force.
Owner:CCCC HIGHWAY BRIDGES NATIONAL ENGINEERING RESEARCH CENTRE CO LTD

System and method with universal segment embeddings for open-vocabulary image segmentation

A computer-implemented system and method relates to open-vocabulary image segmentation. A set of data pairs is automatically generated using a digital image and a corresponding caption. The set of data pairs include image segments and corresponding text data. The set of data pairs includes (i) a first subset that includes object segments as the image segments and corresponding object data as the text data and (ii) a second subset that includes part segments as the image segments and corresponding part data as the text data. A universal segmentation embedding (USE) model includes an image encoder and a segment embedding head. The image encoder generates patch embeddings based on patches of the digital image. The segment embedding head generates segment embeddings based on the image segments and the patch embeddings. Semantic segmentation data is generated based on the segment embeddings.
Owner:ROBERT BOSCH GMBH

Cervical lesion intercellular relation modeling and analysis system based on graph neural network

InactiveCN120747012AImage enhancementMedical data miningCervical lesionCervical tissue
The invention discloses a cervical lesion intercellular relation modeling and analysis system based on a graph neural network, and the system comprises a medical image collection module which is used for collecting a digital image of a cervical tissue pathological section or a cervical TCT slide; the cell detection and segmentation module is used for extracting spatial position information and morphological characteristics of cells; the cell feature extraction module is used for extracting and fusing the spatial position, morphology, texture and biological marker features of the cells; the cell relation graph construction module is used for constructing a heterogeneous cell relation graph with cells as nodes and inter-cell relations as edges; the graph neural network analysis module is used for carrying out feature learning and modeling on the heterogeneous cell relation graph; the intelligent auxiliary diagnosis module is used for generating auxiliary diagnosis suggestions; and the data management and automatic control module is used for realizing automatic control and case data management of the whole process of the data. The intelligent and automatic level of cervical lesion cell analysis can be comprehensively improved, and the accuracy and efficiency of diagnosis are improved.
Owner:HANGZHOU WEIJIN TECHNOLOGY CO LTD

Textile color fastness prediction method based on computer assistance

The invention relates to the technical field of textile detection, in particular to a computer-aided textile color fastness prediction method, which comprises the following steps: acquiring multi-source data before and after a textile color fastness test through computer control equipment, including digital images, reflection spectrum data, process parameters and dye characteristic component data, carrying out pretreatment; extracting color difference features and spectrum similarity features, and screening by adopting a dimension reduction algorithm to form a core feature vector; a machine learning algorithm is adopted to construct a color fastness prediction model, parameters are optimized in combination with cross validation and an early stop mechanism, model fine tuning is executed, and a final optimization model is obtained; and outputting a color fastness grade prediction result, and carrying out spectrum similarity secondary verification. According to the method, objective, efficient and high-precision prediction of the color fastness grade of the textile is realized, and the problems of high subjectivity and low efficiency of traditional manual evaluation are solved.
Owner:JIANGSU BAOMAN BEDROOM ARTICLES

Joint framework for object-centered shadow detection, removal, and synthesis

The present disclosure relates to systems, methods, and non-transitory computer-readable media that detects shadows, removes shadows, and synthesizes shadows in a joint-framework. In particular, the disclosed systems access an object mask of an object and a digital image depicting the object and a shadow of the object. Furthermore, the disclosed systems perform object-centered shadow detection and removal to generate a modified digital image without the shadow by utilizing a shadow analyzer model. Moreover, the disclosed systems receive a user interaction to manipulate an object and generate a modified shadow utilizing a shadow synthesis model where the shadow synthesis model is conditioned on a shadow mask generated by the shadow analyzer model.
Owner:ADOBE INC

Shearing behavior dynamic correction method and system based on wear state recognition

The invention relates to the technical field of image recognition, in particular to a shearing behavior dynamic correction method and system based on wear state recognition. Acquiring a digital image sequence of the cutting edge area; constructing an image feature separation network, and performing parallel feature extraction on the preprocessed digital image sequence; identifying a pixel-level high-frequency texture discontinuous region and an edge gradient direction field by utilizing continuity characteristics of a cutting edge surface periodic texture mode to obtain a target defect probability graph; identifying a projection shadow area and a low-frequency illumination halation of the edge of the bulge by using a backlighting imaging model to obtain an interference artifact probability graph; establishing spatial mutual exclusion constraints of the target defect probability graph and the interference artifact probability graph in a pixel space, and generating a defect binary mask; performing multi-dimensional texture feature calculation on an area corresponding to the defect binary mask, and constructing a surface state feature vector; according to the invention, based on the surface state feature vector, the defect mode category is discriminated, and the corresponding shearing correction parameter is generated.
Owner:SUZHOU LILAI IRON & STEEL CO LTD

System and method for robust inference of heterogeneous material properties via infinite-dimensional integrated digital image correlation

An exemplary system and method that employ inverse-problem analysis that can determine spatially-varying mechanical parameters in a spatially-varying field of a heterogeneous material. Mathematically, the computation simultaneously poses the inversion program and an image registration problem in a continuum limit function space setting to derive a discretization dimension-independent algorithm for the robust inference of heterogeneous material properties. The algorithm can operate using two or more images of a speckled pattern or other non-uniform patterns applied to, or observable of, the surface of the material in a first state and a second state different from the first state. The difference can be used to assess, via a Newtonian-based operator, the infinite-dimensional spatial fields as state variables that are regularized via a regularization model to constrain the inherent ill-posed nature of inverse problems.
Owner:BOARD OF RGT THE UNIV OF TEXAS SYST

Image tampering detection method and device based on cross-modal ViT architecture

The invention provides an image tampering detection method and device based on a cross-modal ViT architecture. The method and device are particularly suitable for positioning digital image tampering operations such as copying-moving, splicing and erasing. The method comprises the following steps: acquiring a to-be-detected image, and inputting the to-be-detected image into a pre-established cross-modal double-flow ViT architecture; the cross-modal double-flow ViT architecture comprises a space domain ViT encoder layer, a frequency domain ViT encoder layer and a deformable cross attention fusion layer; extracting spatial domain feature data corresponding to the to-be-detected image through the spatial domain ViT encoder layer; extracting frequency domain feature data corresponding to the to-be-detected image through a frequency domain ViT encoder layer; and performing dynamic gating fusion on the spatial domain feature data and the frequency domain feature data through a deformable cross attention fusion layer to obtain fusion features, and performing image tampering detection based on the fusion features. According to the method, the problem of long-distance dependence is solved based on a self-attention mechanism of Vision Transform, and picture tampering detection is realized through spatial domain and frequency domain cross-modal conjoint analysis. Technical capability is provided for application scenes such as digital forensics and content auditing, and the material authenticity detection level is improved.
Owner:CHINA UNIONPAY MERCHANT SERVICES CO LTD

Determining outlier images based on category-based image relevance using embedding neural networks

This disclosure describes a framework for determining the category-based image relevance of digital images associated with entities or topics. Specifically, this disclosure describes an image relevance system that determines outlier images within a set of images associated with an entity or topic by correlating semantic content with visual content. For example, the image relevance system ensures that only images relevant to the entity or topic are provided in response to a user query about the entity or topic. The image relevance system can also filter out images from an image set that do not correspond to user input in a search query before providing the image set. Furthermore, the image relevance system can prevent irrelevant images from being added to an image set associated with an entity or topic.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Digital video editing based on a target digital image

Digital video editing techniques are described that are based on a target digital image. In one or more implementations, inputs are received. The inputs include a target text prompt, a target digital image depicting a target object, and a source digital video having a plurality of frames depicting a source object. Regions-of-interest are identified in the plurality of frames of the source digital video, respectively, based on the target text prompt and the target digital image using a machine-learning model, e.g., a diffusion model. A plurality of frames of a target digital video are generated as having the target object using a generative machine-learning model. The generating is based on the regions-of-interest, the target digital image, the source digital video, and a source text prompt describing the source digital video.
Owner:ADOBE INC

Cylindrical roller trajectory tracking method and device

The invention discloses a cylindrical roller trajectory tracking method and device, and the cylindrical roller trajectory tracking method comprises the following steps: obtaining an actual measurement displacement sequence of a key point on the end surface of a roller under a global coordinate system based on a multi-reference matching digital image correlation method; drawing a key point displacement curve about key points and time according to the actually measured displacement sequence; judging the motion state of a roller according to the shape characteristics of the key point displacement curve, dividing a smooth part in the key point displacement curve into a rolling area, and dividing a part with jumping or sharp points into a sliding area; calculating the actual revolution angular velocity and the actual rotation angular velocity of the roller based on the actually measured displacement sequence corresponding to the sliding area; according to the method and the device, the image matching precision can be improved, meanwhile, the rolling and sliding components of the roller are accurately distinguished, the precision of the track of the cylindrical roller is improved, and high-precision and visual experimental data are provided for existing dynamics and tribology models.
Owner:HENAN UNIV OF SCI & TECH

Historical relic image virtual restoration method based on texture reconstruction and color correction

The invention relates to the technical field of image processing and cultural relic protection, in particular to a cultural relic image virtual restoration method based on texture reconstruction and color correction, and the method comprises the following steps: S1, obtaining a digital image of the surface of a to-be-restored cultural relic, and carrying out the denoising and brightness equalization processing of the digital image; s2, identifying and segmenting a damaged area in the image; s3, selecting an optimal sample block which is most matched with the structure of the region to be filled; s4, performing adaptive affine transformation on the optimal sample block to generate a reconstructed texture block; s5, carrying out linear transformation to obtain a repaired texture block after color correction; and S6, seamlessly embedding the repaired texture block after color correction into the damaged area. According to the method, the texture direction alignment and the color distribution correction are combined, so that the consistency of the damaged area and the surrounding image in structure and color is realized, and the naturalness and integrity of cultural relic image restoration are remarkably improved.
Owner:CHONGQING UNIV