Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

35 results about "Visual artifact" patented technology

Visual artifacts (also artefacts) are anomalies apparent during visual representation as in digital graphics and other forms of imagery, especially photography and microscopy.

Proximity-based generation of surface textures for simulated environmental systems and applications

In various examples, a simulation platform generates a simulated driving environment by processing road map data to derive the location of wear-related visual artifacts for sections of a lane surface. Using map data, the simulation platform generates texture maps for aesthetic road rendering, which are used to apply textures to a 3D polygon topology mesh. The simulation platform generates visual artifacts representing the wear and tear of the road surface based on a calculation of one or more lane feature distances derived from the map data. The simulation platform calculates distances associated with reference line data derived from an image to render a texture of one or more lane features.The spacing is used to determine how the appearance of the texel is adjusted to include wear-related visual artifacts when displayed on a lane of the simulated driving environment.
Owner:NVIDIA CORP

Proximity-based surface texture generation for simulated environment systems and applications

Proximity-based surface texture generation for simulated environment systems and applications is disclosed. In various examples, a simulation platform generates a simulated driving environment by processing road map data to infer locations of wear-related visual artifacts for portions of a roadway surface. The simulation platform uses the map data to generate texture maps for aesthetic road rendering, which are used to apply texture to the 3D polygon topology mesh. The simulation platform generates visual artifacts representing use and wear of the roadway surface based on calculating one or more distances from roadway lane features derived from the map data. The simulation platform calculates a distance from the one or more roadway lane features associated with reference line data derived from the image to render the texture. The distances are used to determine how to adjust the appearance of texels to include wear-related visual artifacts when rendering wear-related virtual artifacts on the roadway of the simulated driving environment.
Owner:NVIDIA CORP

Lightweight diffusion virtual fitting algorithm for high-resolution clothing human body image

The invention discloses a lightweight diffusion virtual fitting algorithm for a high-resolution clothing human body image. The algorithm specifically comprises the following steps: step 1, acquiring a disclosed high-resolution virtual fitting data set; 2, preprocessing the clothing image and the reference figure image in the high-resolution virtual fitting data set obtained in the step 1 to obtain a figure image irrelevant to clothing; 3, constructing a lightweight diffusion virtual fitting model aiming at the high-resolution clothing human body image, and designing a required loss function; and step 4, training the lightweight diffusion virtual fitting model for the high-resolution clothing human body image by using the high-resolution clothing human body data set in the step 1. According to the method, the problems of detail loss and visual artifacts in a high-resolution scene in a virtual fitting method based on the generative adversarial network are solved, and the defects that a traditional diffusion model is huge in volume, high in computing resource demand and difficult to retain local features are overcome.
Owner:XI'AN POLYTECHNIC UNIVERSITY

Barcode scanning image enhancement method and system based on deep learning

The invention discloses a bar code scanning image enhancement method and system based on deep learning, and relates to the technical field of deep learning and image processing, and the method comprises the steps: collecting a bar code image, carrying out the preprocessing, and generating a standardized and grayscale image; constructing a difference value and a gradient response diagram based on the grayscale image, obtaining an attention map according to the standardized image, performing linear superposition in combination with the difference value and the gradient response diagram, obtaining a combined attention map, performing masking and shielding, and generating a mask and a shielding map; carrying out image interception on the mask image and the joint attention image based on the mask image and the calculation external rectangular frame, then executing channel splicing, obtaining an input tensor, inputting the input tensor into a conditional generative adversarial network, and outputting a repaired image; according to the invention, natural transition between a restoration result and an original image is realized by using a pixel-level fusion strategy based on Gaussian weight, and common visual artifacts such as edge fault and brightness jump can be effectively avoided.
Owner:HANGZHOU XIANGHE ZHENCAI TECH CO LTD

Photometric image enhancement for endoscopy

This disclosure provides methods, devices, and systems for navigating medical instruments. The present implementations more specifically relate to photometric image enhancement techniques for endoscopy. In some aspects, a machine learning model may be trained to infer an enhanced image from a low-quality image captured by the camera of an endoscope. As used herein, the term “low-quality image” refers to any image containing visual artifacts, obstructions, and / or other deficiencies. By contrast, an “enhanced image” is a digitally modified representation of a low-quality image that removes and / or corrects at least some of the visual artifacts, obstructions, or other deficiencies in the low-quality image. A controller for a medical system may extract information from the enhanced image based on one or more image processing operations and generate a graphical user interface (GUI) for navigating the instrument within the anatomy based at least in part on the information extracted from the enhanced image data.
Owner:AURIS HEALTH INC

An image restoration method based on a hybrid structured sparse model

ActiveCN116452443BImprove image restoration effectImage enhancementImage analysisPattern recognitionSparse model
This invention discloses an image restoration method based on a hybrid structured sparse model, specifically including the following steps: initializing the restored image and setting the number of iterations; constructing a matrix of similar image patch groups; establishing a hybrid structured sparse model; using the hybrid structured sparse model to sparsely encode each similar image patch group; reconstructing each similar image patch group; and restoring the entire image based on all reconstructed similar image patch groups. The image restoration method of this invention better reconstructs details such as edges and textures, and effectively suppresses unwanted visual artifacts, further improving the image restoration effect.
Owner:XIAN UNIV OF TECH

Remote sensing image automatic registration method, device, equipment, medium and program product

The invention relates to a remote sensing image automatic registration method and device, equipment, a medium and a program product. The method comprises the steps of determining an effective overlapping region and a non-overlapping region of a reference image and a to-be-registered image, adaptively selecting a global registration process or a block registration process according to the size of the effective overlapping region, and obtaining an image pair subjected to corresponding registration processing; performing screening to obtain feature matching point pairs meeting a preset confidence threshold; according to the feature matching point pairs, respectively executing geometric transformation estimation and image correction of an adaptive matching corresponding registration process on the effective overlapping region and the non-overlapping region; and under the condition that the block registration process is executed, intelligent splicing is executed through a global optimal seam fusion algorithm based on dynamic programming, and a second registration result image with visual seamless is generated. By adopting the method, the feature matching robustness and precision in a complex scene can be improved, visual artifacts of image splicing after registration are effectively eliminated, and visual seamless fusion is realized.
Owner:SHANG HAI ZHANG JIANG SHU XUE YAN JIU YUAN

Wearable heads-up displays including combiner with visual artifact reduction

PendingUS20260153738A1Optical elementsComputer graphics (images)Dichroic prism
A wearable heads-up display (WHUD) reduces the prevalence of visual artifacts by employing a projector with a combiner including a plurality of combiner elements such as dichroic prisms. The projector is configured with one or more features that 1) reduce the amount of stray light generated at the combiner 2) change the path of the stray light so that the stray light is unable to exit the projector and thus is unable to create visual artifacts, or any combination thereof. By reducing the stray light that is generated and by changing the path of the stray light as described herein, the likelihood that a user will see a visual artifact is reduced, thus improving the user experience with the WHUD system.
Owner:GOOGLE LLC

A general robust reinforcement method and system for a circuit board solder joint defect recognition model

PendingCN122289761APattern recognitionAlgorithm
This invention discloses a general robustness enhancement method for a circuit board solder joint defect identification model. First, training samples are input into a pre-trained general diffusion generation model to extract latent variables. A flexible penalty term based on ReLU properties is used to replace hard truncation to constrain the perturbation range, leveraging the strong structural prior of the diffusion model to ensure that the generated latent space adversarial samples highly conform to the distribution of real solder joint images. Second, a smoothing period is introduced during the adversarial optimization iteration of latent variables, and a strategy based on historical averages is used to periodically update the optimization trajectory to filter out high-frequency components, thereby avoiding visual artifacts caused by drastic fluctuations in latent variables and interference with the learning of real features. Finally, before calculating the adversarial loss, the generated samples are subjected to multi-scale scaling and multi-directional translation, and the target model weights are randomly zeroed multiple times to simulate the training blind zone, so that the calculation of the adversarial gradient covers the uncertainty of spatial geometric transformation and model dynamic evolution.
Owner:HUNAN KUANGAN NETWORK TECH CO LTD

Light emitting display with tiles and data processing

A light emitting display can be formed from tiles mounted within a certain distance range with respect to each other and with an established blending region positioned towards the edges of the tiles. A tile can be a matrix of light emitting elements, such as LEDs, OLEDs, quantum dots, or other element that emits light. The tolerance of spacing between tiles can allow for less precision in alignment during installation in a theatre, thereby reducing display assembly cost but still maintaining a display for displaying an image at a high quality with reduced or eliminated appearance of visual artifacts between tiles.
Owner:IMAX CORP

Computer-implemented method for reviewing at least one determined organ contour in medical imaging data

PendingUS20260080532A1Image enhancementImage analysisMedical imaging dataVisual artifact
A method according to an example embodiment includes determining at least one organ contour of medical imaging data, the medical imaging data comprise at least one image of at least a part of at least one organ; at least one of, determining an image uncertainty information describing at least one area of the at least one image with a visual artifact, determining an organ contour inaccuracy information describing if the determination of the at least one organ contour is inaccurate, or determining a warning information describing at least one of a predetermined warning concerning the organ or an area surrounding the organ for which the organ contour was determined; and determining review data comprising the at least one image and the determined at least one organ contour and at least one of the determined image uncertainty information, the determined organ contour inaccuracy information or the determined warning information.
Owner:SIEMENS HEALTHINEERS AG

Image super-resolution reconstruction method based on multi-scale mixed attention residual network

The invention discloses an image super-resolution reconstruction method based on a multi-scale mixed attention residual network, and belongs to the technical field of computer vision and image processing. The invention provides an end-to-end reconstruction network (MHARN) aiming at the problems of insufficient detail recovery capability, limited cross-dimension feature interaction and poor adaptability to different image contents in the prior art. The method comprises the following steps: firstly, dynamically aggregating multi-scale features through parallel convolution branches with different voidage by utilizing a progressive convolution group (PCG), and breaking through the limitation of a single receptive field; the features are input into a cascaded enhanced residual attention block (ERAB), a multi-head hybrid attention module (MHAM) in the ERAB is utilized to capture window texture, geometric structure and global channel information in parallel, and a dynamic feature enhancement module (DFEM) is combined to adaptively generate convolution kernel parameters according to local image content; and finally, a high-resolution image is obtained through global feature fusion and up-sampling reconstruction. The texture detail recovery capability can be effectively enhanced, the visual artifacts are remarkably reduced, and the image reconstruction quality is improved.
Owner:WUHAN CITY VOCATIONAL COLLEGE

Generating virtual representations using media assets

Generating the 3D representation of the object includes obtaining sensor data of the object. Media assets including objects may be obtained from a digital asset repository. Visual artifacts of an object may be generated from a media asset and used with sensor data to generate one or more virtual representations of the object. The visual artifacts include texture and / or geometric characteristics of the visual appearance of the object and are derived from image data in the media asset.
Owner:APPLE INC

Proximity-based surface texture generation for simulated environment systems and applications

In various examples, a simulation platform generates a simulated driving environment by processing road map data to infer the location of wear-related visual artifacts for portions of a roadway surface. Using map data, the simulation platform generates texture maps for aesthetic road renderings that are used to apply textures onto a 3D polygon topology mesh. The simulation platform generates visual artifacts representing use and wear of the roadway surface based on calculating one or more distances from roadway lane features derived from the map data. The simulation platform computes distances associated with reference line data derived from an image to render texture from one or more roadway lane features. The distances are used in determining how the appearance of the texels are adjusted to include the wear-related visual artifacts when rendered on a roadway of the simulated driving environment.
Owner:NVIDIA CORP

Adaptive encoding based on individual game player's sensitivity to visual artifacts

Techniques are described for an encoder or decoder to adaptively change (e.g., 604) code processing based on a particular user's sensitivity to flicker or flash or blocking or other visual artifacts. Alternatively, the video may be pre-processed (e.g., 1000) to suppress artifacts based on the sensitivity of the user prior to encoding, and / or post-processed to suppress artifacts after decoding.
Owner:SONY INTERACTIVE ENTERTAINMENT LLC

Methods and apparatus for encoding, decoding, and rendering 6DOF content from 3DOF+ components.

This invention discloses a method where volumetric content is encoded into a cluster by an encoder and transmitted to a decoder that retrieves the volumetric content. Clusters shared across different viewpoints are obtained and collectively managed. These clusters are then projected onto a 2D image and encoded into independent video streams. This achieves reduction in visual artifacts, as well as reduction in storage and streaming data.
Owner:INTERDIGITAL VC HOLDINGS INC

Dynamically generating user interfaces based on machine learning models

Certain aspects of the present disclosure provide techniques for rendering visual artifacts in virtual worlds using machine learning models. An example method generally includes identifying, based on a machine learning model and a streaming natural language input, an intent associated with the streaming natural language input; generating, based on the identified intent associated with the streaming natural language input, one or more virtual objects for rendering in a virtual environment displayed on one or more displays of an electronic device; and rendering the generated one or more virtual objects in the virtual environment.
Owner:INTUIT INC

A speech video generation method based on audio and video structure alignment

The application discloses a speech video generation method based on audio and video structure alignment and belongs to the virtual digital person field.The application comprises an audio segmentation module, an audio conversion module, an audio coding module, a video coding module and a video fusion decoding module.The audio conversion module is used for converting segmented phonemes into a mel-frequency spectrogram which is more in line with the frequency range of human ears according to Fourier transform.In the audio coding process, the frames of the same phonemes are taken as a continuous time module, and the time module is taken as time consistency to constrain the change of the lips, so that the fine-grained control of the lips is realized at the phoneme level through the time consistency constraint.In the video coding, the part region of the multi-pose change face in the input video is set as a mask region, the mask region is used as spatial consistency to accurately control the change amplitude of the lips, the position of the speaker's lips is aligned, the visual artifacts of the video are reduced, the facial details are optimized, and a high-quality speech video with audio-visual synchronization is generated.
Owner:BEIJING INST OF TECH

Neural-Network Based Visibility Modeling Using Just Noticeable Difference Training Data

A neural-network is trained to produce a visibility model based on just noticeable difference training data. The trained network is used to generate visibility masks to control the visual quality of image processing operations, such as digital watermarking, compression and other image editing operations. The visibility model is refined by iteratively adjusting for over and under predictions of a just noticeable difference threshold via an interactive user interface that pinpoints visual artifacts and captures updates from users, who provide adjustments such that the artifact is below what they perceive to be a just noticeable difference. Image processing, such as digital watermarking is performed in real time to provide the user with feedback to ascertain just noticeable difference levels corresponding to watermark embedding operations.
Owner:DIGIMARC CORP

Remote sensing image automatic registration method, device, equipment, medium and program product

This application relates to a method, apparatus, device, medium, and program product for automatic registration of remote sensing images. The method includes: determining the effective overlapping and non-overlapping regions of a reference image and an image to be registered; adaptively selecting a global registration process or a block registration process based on the size of the effective overlapping region to obtain corresponding image pairs for registration processing; filtering feature matching point pairs that meet a preset confidence threshold; performing geometric transformation estimation and image correction adapted to the corresponding registration process on the effective overlapping and non-overlapping regions based on the feature matching point pairs; and, in the case of executing the block registration process, performing intelligent stitching through a globally optimal seam fusion algorithm based on dynamic programming to generate a visually seamless second registration result image. This method can improve the robustness and accuracy of feature matching in complex scenes, effectively eliminate visual artifacts in the stitching of registered images, and achieve visually seamless fusion.
Owner:SHANG HAI ZHANG JIANG SHU XUE YAN JIU YUAN

Image processing method and device and storage medium

The invention provides an image processing method and device and a storage medium, and the method comprises the steps: carrying out the depth estimation of an initial image, and obtaining an initial depth map; performing expansion processing based on the initial depth map to obtain a target depth map texture; and rendering based on the viewpoint offset determined by the interaction event, the target depth map texture and the initial image to obtain a target rendering frame. According to the method, the initial depth map obtained by estimating the depth of the initial image is subjected to expansion processing, so that one or more pixels are expanded from a region with brighter pixels to a region with darker pixels, and a depth region slightly larger than the visual contour of an object is formed; and then, rendering is performed by using the viewpoint offset, the target depth map texture and the initial image to determine a target rendering frame, so that the problem of visual artifacts during rendering can be avoided.
Owner:HUNAN HAPPLY SUNSHINE INTERACTIVE ENTERTAINMENT MEDIA CO LTD

Image recognition-based decoration engineering construction quality detection method and system

This application proposes a method and system for inspecting the construction quality of decorative engineering based on image recognition, belonging to the field of image recognition technology. The method includes acquiring visible light images, three-dimensional morphological data, and thermophysical state data of the decorative material to be inspected. The visible light images characterize the visual features of the decorative material, the three-dimensional morphological data characterizes the three-dimensional structural features, and the thermophysical state data characterizes the surface temperature characteristics. The visible light images are analyzed to identify potential abnormal areas on the decorative material. Based on the visible light images, three-dimensional morphological data, and thermophysical state data, the potential abnormal areas are judged to distinguish between real defects and visual artifacts, resulting in a judgment result. Based on the judgment result, a report containing confirmed real defect information is generated. This application can improve the accuracy and reliability of inspection.
Owner:SHANGHAI YIJI ARCHITECTURAL DECORATION ENG CO LTD

Generalized moving text region detection for broken text recovery

One embodiment provides a method comprising receiving content for presentation on a display, and obtaining a plurality of image patches from the content. Each image patch is segmented from a bottom and horizontal-center portion of a sub-sampled grayscale image of a frame of the content. The method further comprises applying a searching algorithm to the plurality of image patches to detect a region of interest of moving text in the content. The searching algorithm utilizes a first set of features and a second set of features to detect the region of interest with respect to a vertical axis and a horizontal axis of the display, respectively. The method further comprises correcting one or more visual artifacts in the region of interest, where the one or more visual artifacts include broken text.
Owner:SAMSUNG ELECTRONICS CO LTD

Ophthalmic devices comprising light stable mimics of macular pigments and other visible light filters

An ophthalmic device containing a visible light filter is described. The present invention provides an ophthalmic device that is the product of a free radical reaction of a reactive mixture comprising: one or more monomers suitable for manufacturing ophthalmic devices; a first visible light filter compound having a maximum visible light absorption between 430 nm and 480 nm and a full width at half maximum (FWHM) of at least 35 nm and at most 150 nm at the maximum visible light absorption, wherein the first visible light filter compound is photostable, and wherein the first visible light filter compound has a concentration of at least 7740 L·mol⁻¹. ‑1 .cm ‑1 The device contains a molar extinction coefficient and a second visible light filter compound having a maximum visible light absorption between 480 nm and 530 nm and a full width at half maximum (FWHM) of at least 50 nm and at most 150 nm. The reactive monomer mixture may also contain a third visible light filter compound. These devices can provide one or more visually beneficial effects, including enhanced macular pigment optical density, improved color perception and color enhancement, and reduced visual artifacts.
Owner:JOHNSON & JOHNSON VISION CARE INC

Dynamic path planning methods for large-scale raster maps

ActiveCN121163531BAlgorithmVisual artifact
This invention relates to the field of path planning technology, and provides a dynamic path planning method for large-scale raster maps. It constructs a three-layer policy mask comprising a permitted normal layer, an execution-state temporary extension layer, and a service alarm layer to clearly delineate path regions with different functions. Initial path planning is performed at the permitted normal layer based on a region of interest-thick line view-through direct pull search strategy, ensuring the normal security and stability of the initial path. Detour path planning is performed at both the permitted normal layer and the execution-state temporary extension layer based on the same strategy, enabling real-time, efficient, and secure orderly planning when detour requirements are detected. Regression convergence processing is performed on candidate detour path segments based on the baseline path and the initial rendered path, further improving path security and smoothness. Corner alignment resampling processing is performed on intermediate rendered paths to effectively avoid visual artifacts.
Owner:TUS DIGITAL DISPLAY TECH (SHENZHEN) CO LTD +1

Method for optimizing live cell clathrin imaging parameters based on srrf nanoscope and related apparatus

The application discloses a kind of live cell clathrin imaging parameter optimization method and related device based on SRRF nanoscope, it is related to biomedical imaging field, the method includes under live cell culture condition, when using TIRF microscope to the time series imaging of live cell sample expressing fluorescent label, original image sequence is obtained;It is input NanoJ-SRRF processing system, a plurality of core reconstruction parameters are systematically combined, and several super-resolution reconstruction images are generated;For the super-resolution image reconstructed by each combination parameter, comprehensive quantitative evaluation is carried out using multimodal evaluation system, and multimodal quantitative quality evaluation result is obtained;Based on this result, the parameter combination that meets the preset resolution, structure fidelity index and the minimum visual artifact is screened out, and the optimal parameter set for the imaging of live cell clathrin coated pit is determined.The application can reduce reconstruction artifact, realize high-fidelity nanoscale observation of live cell CCP dynamics.
Owner:ZHEJIANG UNIV

3D Gaussian reconstruction optimization method based on display depth partition

The invention discloses a 3D Gaussian reconstruction optimization method based on display depth partitioning, and belongs to the technical field of computer graphics and three-dimensional display. The method comprises the following steps: firstly, constructing a core area and an edge area by taking a display focal plane as a center according to an effective depth-of-field range of three-dimensional display equipment, and establishing a double-depth buffer area; the marginal region is then subdivided into a near marginal region and a far marginal region. Executing all-parameter optimization without the upper limit of the number of iterations in the core area and adopting high-density control; introducing a depth weight-based clone splitting probability attenuation mechanism and a spatial scale constraint into the near edge region; and applying the attenuation mechanism in the far edge region and applying the maximum iteration number limitation, and finally outputting a three-dimensional Gaussian scene model of partition optimization. According to the method, through differentiated resource allocation, the calculation efficiency is remarkably improved, the visual artifacts at the boundary of the depth of field are effectively eliminated, and the equipment suitability is enhanced.
Owner:SHENZHEN RES INST OF BEIJING UNIV OF POSTS & TELECOMM

Cloud edge-end cooperative monitoring video reconstruction method, system and device based on large model semantic-motion priori guidance and medium

The invention relates to a cloud side end cooperative monitoring video reconstruction method, system, equipment and medium, in particular to a cloud side end cooperative monitoring video reconstruction method, system, equipment and medium based on large model semantics-motion priori guidance. The invention aims to solve the problems of key semantic information loss, error accumulation, visual artifacts, system efficiency and deployment bottleneck in the prior art. According to the method, lightweight prior is intelligently extracted and transmitted on the edge side, a traditional mode of periodically transmitting high-resolution key frames is replaced, so that the transmission bandwidth can be greatly reduced, rich knowledge provided by a large model is utilized to accurately guide the reconstruction process, error accumulation and artifacts are effectively inhibited, and the reconstruction efficiency is improved. And high-fidelity and high-stability recovery of the monitoring video is realized. The invention belongs to the technical field of video reconstruction.
Owner:HARBIN INST OF TECH

An interpretable text image tampering detection method and system based on visual artifact and semantic anomaly fusion

ActiveCN121236781BMake up for the black box phenomenon that cannot explain the detection principleImprove generalization abilityInstrumentsPattern recognitionLinguistic model
The application provides an explainable text image tampering detection method and system based on fusion of visual artifacts and semantic anomalies, which comprises the following steps: obtaining a tampered text image and its corresponding tampering analysis report pair for data preprocessing; constructing a tampering locator and training the tampering locator based on the data preprocessed tampered text image and tampering analysis report; constructing a tampering interpreter and training the tampering interpreter based on the data preprocessed tampered text image and tampering analysis report; obtaining a tampering detection model based on the trained tampering locator and tampering interpreter, and inputting a to-be-detected image into the tampering detection model to realize tampering detection. The application utilizes a visual expert model and a multimodal large language model to work cooperatively, realizes accurate positioning and cause explanation of the tampering area, and maintains high detection accuracy and robustness under compression, blurring and other degradation and cross-scene conditions, and has high practicability.
Owner:SOUTH CHINA UNIV OF TECH

Systems and Methods for On-Cell Touch Off-State Pattern Visibility Mitigation

An electronic display may include both a display subsystem and a touch subsystem. Opaque metal layers used in on-cell touch sensor technology of the touch subsystem may cause undesirable visual artifacts. Visual artifacts caused by the cuts in the touch metal mesh may be reduced or eliminated by disposing a metal patch on a different layer than the metal mesh to increase reflected light, presenting a more uniform appearance to the user. Visual artifacts caused by a bridge disposed across the metal mesh may be reduced or eliminated by curving or angling the geometry of the bridge across the metal mesh or by disposing a metal cladding (e.g., covering) above the bridge. Visual artifacts caused by functional vias disposed in a net of the touch subsystem may be reduced or eliminated by disposing non-functional vias and / or non-functional holes in the net.
Owner:APPLE INC