Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

5650results about "Image coding" patented technology

BIM (Building Information Modeling) intelligent management platform and method for project construction full life cycle

The invention provides a BIM intelligent management platform oriented to a whole life cycle of project construction. A building information model, Internet of Things sensing data and a block chain evidence storage mechanism are integrated through a multi-source data fusion technology, and a whole-process data chain of association planning, design, construction, operation and maintenance is associated. The platform adopts space optimization Huffman coding to realize model lightweight, combines a constraint genetic algorithm to optimize a construction path, and applies a bidirectional long-short-term memory network to analyze an equipment state. A three-chain block chain system is reconstructed on the architecture, intelligent association of engineering quantity and payment nodes is realized through cooperation of a main chain, a calculation quantity side chain and an auditing side chain, and mobile terminal offline interaction is supported based on a digital-analog separation technology. The platform covers an intelligent design management unit, a block chain investment management unit, a dynamic correction management unit, a quality safety responsibility tracing unit, an NLP risk management unit and a digital twin operation and maintenance unit. The units achieve cross-system cooperation through a unified data bus, and a closed-loop management architecture covering the whole life cycle of project construction is formed.
Owner:DONGGUAN DAYE CONSTRUCTION TECHNOLOGY CONSULTING CO LTD +1

Variable-bit-rate image compression method and system, apparatus, terminal, and storage medium

The present disclosure provides a variable-bit-rate image compression method and system, an apparatus, a terminal, and a storage medium. The variable-bit-rate image compression method includes: obtaining an initial feature map from a to-be-encoded image; quantizing the initial feature map by a dead-zone quantizer; performing entropy encoding on the quantized feature map and hyper-prior information to obtain a compressed bit-stream; performing entropy decoding on the compressed bit-stream, and recovering quantized hyper-prior information and the quantized feature map; performing inverse quantization on the quantized feature map to obtain a reconstructed feature map; obtaining a reconstructed image from the reconstructed feature map; and adjusting quantization and inverse quantization parameters according to a target bit-rate or target distortion. The present disclosure provides a precise bit-rate control solution, makes the bit-rate of the compressed bit-stream better adapt to the dynamic change of a network bandwidth, and has an extremely high actual application value.
Owner:SHANGHAI JIAOTONG UNIV

Dynamic Latent Space Adaptation Based on Spatiotemporal Kernal Context for Multiscale Rendering

A system for dynamic latent space adaptation using spatiotemporal kernel context for multiscale rendering with hierarchical and Lorentzian autoencoders. The Spatiotemporal Kernel Estimator (SKE) analyzes media through motion field, temporal recurrence, frequency band, and scene semantics analyzers to generate adaptive kernel parameters encoding content-specific importance distributions. The system dynamically adapts latent manifold geometry by modifying metric tensor properties according to kernel context, enabling content-aware compression that allocates representational capacity based on visual significance. A multiscale cache implements kernel-adaptive retention policies prioritizing important regions. An adaptive renderer provides intelligent level-of-detail selection based on zoom level and kernel-estimated importance, optimizing processing allocation. The self-optimizing architecture continuously refines kernel context and geometric adaptation based on user interaction and performance feedback, achieving superior compression ratios and perceptual quality. Applications include bandwidth-efficient video streaming, virtual reality, scientific visualization, and cognitive video analytics requiring intelligent context-aware visual processing.
Owner:ATOMBEAM TECH INC

Multi-scale semantic guidance image compression method and system and storage medium

The invention discloses a multi-scale semantic guidance image compression method and system and a storage medium, and the method comprises the following steps: obtaining input image data, carrying out the preprocessing of an input image, and obtaining standardized image data; inputting the standardized image data into a pre-trained semantic segmentation network to generate a multi-scale semantic feature map and a semantic weight map corresponding to the multi-scale semantic feature map; a three-stage pyramid encoder is constructed, and the standardized image data is subjected to the following steps of: sampling under depth separable convolution to generate multi-scale features; the reversible neural network carries out nonlinear transformation on the multi-scale features; the multi-scale feature subjected to nonlinear transformation is decomposed into a low-frequency sub-band and a high-frequency sub-band through adaptive discrete wavelet transformation, dynamic selective state space modeling is executed on the high-frequency sub-band based on a semantic weight map, and a compressed code stream is generated; and inputting the compressed code stream into a decoder, decoding based on a lightweight Mama module, and reconstructing an image in combination with inverse wavelet transform and a semantic weight map.
Owner:XIANGJIANG LAB

Bad weather image restoration method based on multi-modal state space model

The invention discloses a bad weather image restoration method based on a multi-modal state space model, and belongs to the technical field of computer vision. In order to solve the problem that an existing unified bad weather image restoration method has limitations in the aspects of global receptive field and computational efficiency, the unified restoration of various bad weather images is realized by designing a parallel multi-mode encoder and an adapter to generate comprehensive prompts containing degeneration semantics. Through parallel connection of a state space module and a local context sensing module of a double-attention mechanism, simultaneous capture of global long-range dependence and local detailed features is realized. Experiments show that the method is high in generalization ability, high in processing speed and low in calculation complexity, and can show good adaptive capacity on a plurality of bad weather image removal tasks.
Owner:TAIYUAN UNIVERSITY OF TECHNOLOGY

Thyroid ultrasonic robot automatic scanning method, device and equipment based on RGB image and depth information and medium

The invention relates to the technical field of computer vision, and discloses a thyroid ultrasonic robot automatic scanning method, device and equipment based on RGB images and depth information and a medium, and the method comprises the steps: obtaining image information and depth information, coding the depth information, fusing and recognizing a target scanning area, determining an initial scanning point and an initial scanning direction, controlling the scanning probe to scan and collect a real-time scanning image, analyzing the real-time scanning image to recognize a preset target and an artifact area, adjusting a scanning posture and a scanning path based on a recognition result, monitoring a continuous existence state of the preset target, and stopping scanning when the preset target is not recognized continuously. The target area is identified by fusing the multi-modal image information, the scanning posture and path are dynamically adjusted in combination with real-time image analysis, scanning termination is intelligently controlled according to the target detection result, the positioning accuracy, image quality and standardization level of ultrasonic scanning are improved, and the method is suitable for automatic ultrasonic imaging of thyroid and superficial organs.
Owner:SHENZHEN BEAUTIFUL RUBIKS CUBE ROBOT CO LTD

Point cloud data transmission device, point cloud data transmission method, point cloud data reception device, and point cloud data reception method

A point cloud data transmission method according to embodiments may comprise the steps of: encoding point cloud data; and transmitting a bitstream comprising the point cloud data. A point cloud data reception method according to embodiments may comprise the steps of: receiving a bitstream comprising point cloud data; and decoding the point cloud data.
Owner:LG ELECTRONICS INC

Processing method and system for sparse compression reconstruction of LDI sub-pixel image and application

The invention provides a processing method and system for reconstructing an LDI sub-pixel image through sparse compression and application, and the method comprises the steps: calculating a low-resolution to-be-exposed image corresponding to each phase structure in advance through a computer image compression algorithm; the low-resolution to-be-exposed images obtained through calculation are loaded to the digital micro-reflector according to the time sequence, and the display moment of each pattern is matched with the rotation position of the corresponding phase structure; coupling the pattern of the digital micro-mirror and the phase encoding mask in a frequency domain through a 4-f optical system, and reconstructing a high-resolution exposure pattern at a sub-pixel level by utilizing a diffraction effect; projecting a high-resolution light field of the reconstructed high-resolution exposure pattern to the surface of the photoresist, and forming a target circuit pattern by accumulating exposure dose; and the micro-nano structure with sub-pixel precision is obtained after development. The system comprises an image compression module, a pattern matching module and a pattern forming module. According to the invention, the manufacturing cost of the laser direct imaging equipment is reduced under the condition of the same precision.
Owner:高峰

Method and apparatus of encoding / decoding point cloud geometry data captured by a spinning sensors head

There is provided methods and apparatus of encoding / decoding a point cloud representing a physical object. Points are captured by a spinning sensors head and are represented by sensor indices associated with sensors that captured the points, azimuthal angles representing capture angles of said sensors, and radius values of spherical coordinates of the point. Points are ordered based order indices obtained from the azimuthal angles and the sensor indices. Order index differences are encoded. An order index difference represents a difference between order indices associated with two consecutive ordered points. Optionally, the method encodes radius values, residual azimuthal angles associated with ordered points and residuals of three-dimensional cartesian coordinates of ordered points based on their three-dimensional cartesian coordinates, decoded azimuthal angles based on azimuthal angles, decoded radius values and sensor indices.
Owner:BEIJING XIAOMI MOBILE SOFTWARE CO LTD

Method and system for generating a 3D parametric mesh of an anatomical structure

There is provided a method and a system for generating a 3D parametric mesh of an anatomical structure of a patient for storing multi-domain data therein. A plurality of anatomical segments having been obtained from segmentation of a set of images of a patient are received. A 3D mesh comprising a plurality of concentric 3D mesh layers is received, where each concentric 3D mesh layer includes a same predetermined number of nodes. A set of nodes in the 3D mesh corresponding to a respective anatomical segment is determined to obtain a respective correspondence rule therebetween. The set of nodes is encoded with a set of features from the respective anatomical segment by using the correspondence rule to obtain a 3D parametric mesh, each node of the set of nodes in the 3D parametric mesh being associated with a respective plurality of feature channels comprises the set of features.
Owner:VITAA MEDICAL SOLUTIONS INC

Method and system for predicting abdominal aortic aneurysm (AAA) growth

There are provided methods, systems and non-transitory storage mediums for predicting growth of an abdominal aortic aneurysm (AAA) of a patient having been diagnosed with AAA. Segmented regions of interest (ROI) comprising the aorta and adjacent structures are received by segmenting a set of images. A wall shear stress parameter and intraluminal thickness parameter is determined. A 3D parametric mesh comprising a plurality of concentric 3D mesh layers is generated, where each concentric 3D mesh layer includes a same predetermined number of nodes. The generation includes encoding the segmented ROIs, the wall shear stress parameter and the intraluminal thickness parameter as features at respective node locations in the 3D parametric mesh. A trained growth prediction machine learning model predicts, based at least on a subset of features of the 3D parametric mesh, if the given patient will show AAA growth. The training of the growth prediction model is also disclosed.
Owner:VITAA MEDICAL SOLUTIONS INC

Cross-source data three-dimensional reconstruction method and system based on improved Gaussian sputtering

The invention discloses a cross-source data three-dimensional reconstruction method and system based on improved Gaussian sputtering, and the method comprises the steps: collecting an unmanned plane inclined image and a ground panoramic image of a target region, and constructing a time-space correlation data set; based on multi-view geometric constraints, space-time coding matching point pairs are established through an adaptive feature pyramid, intelligent incremental cross-source data sparse reconstruction is carried out, and point cloud and camera parameters are output; adopting improved Gaussian sputtering, compressing a three-dimensional Gaussian kernel into a two-dimensional Gaussian primitive through double tangent vector constraint, and fitting surface geometry to realize multi-scale reconstruction; and optimizing primitive parameters by using a differentiatable renderer, completing multi-scale fine reconstruction through gradient back propagation, and generating a high-precision three-dimensional model. According to the method, multi-scale accurate geometric prior input and accurate camera poses are provided for three-dimensional reconstruction, the dependence on professional manual operation in a traditional three-dimensional reconstruction method is greatly reduced, and meanwhile, the geometric accuracy and visual fidelity of a reconstruction result are remarkably improved.
Owner:HANGZHOU INST FOR ADVANCED STUDY UCAS

Multi-view three-dimensional Gaussian densification method and system for adaptive density control

The invention belongs to the technical field of three-dimensional scene reconstruction, and particularly discloses a multi-view three-dimensional Gaussian densification method and system for adaptive density control, and the method comprises the following steps: collecting a multi-view original image, and carrying out the preprocessing of the multi-view original image; complexity features are extracted, a pixel-level complexity heat map is generated, and a globally unified three-dimensional complexity field is constructed; performing back projection on the reconstruction residual error, high-frequency inconsistency and depth / geometric consistency cost of each view angle, generating three-dimensional error popularity, determining a candidate newly-added set and a candidate pruned set, generating a weak label to train a lightweight multilayer perceptron classifier, outputting a ternary probability corresponding to newly-added / pruned / maintained, and obtaining a new / pruned / maintained three-dimensional perceptron classifier; and performing Gaussian densification operation on the newly added region. By adopting the technical scheme, fine point adding is carried out on the complex area, effective pruning is carried out on the simple area, and meanwhile, the synthesis quality, the global consistency and the calculation efficiency of the new view angle are improved.
Owner:CHONGQING UNIV

Personalized image generation using combined image features

Examples described herein relate to personalized image generation using combined image features. A plurality of input images is provided by a user of an interaction application. Each of the plurality of input images depicts at least part of a subject. Each input image is encoded to obtain an identity representation. The identity representations obtained from the plurality of input images are combined to obtain a combined identity representation associated with the subject. A personalized output image is generated via a generative machine learning model. The generative machine learning model processes the combined identity representation and at least one additional image generation control to generate the personalized output image. At a user device, the personalized output image is presented in a user interface of the interaction application.
Owner:SNAP INC

Dynamic mesh geometry refinement component adaptive coding

Computer-implemented methods and systems for processing geometry replacements are disclosed. The methods include decoding / encoding a syntax element associated with a coding mode from / into a bitstream associated with geometry displacements; and reconstructing / converting, based on a coefficient configuration associated with the coding mode, a plurality of quantized transform coefficients from / to a plurality of zero-run length codes.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

Visual encoding and decoding of 3D gaussian splats

An encoder projects a scene represented by 3D gaussian splats into 2D representation(s), where the 2D representation(s) store different parameters of the 3D gaussian splats, along with associated metadata describing transformation from 3D space into 2D representations. The encoder forms the 2D representation(s) and the associated metadata into bitstream(s). The encoder outputs the bitstreams. The decoder receives the bitstream(s) encoding 2D representation(s) and associated metadata. The decoder applies decoding to corresponding individual bitstreams to form the 2D representation(s) and the associated metadata. The decoder reprojects, as described by the associated metadata, the 2D representations) into the scene represented by the 3D gaussian splats. The decoder outputs the scene for rendering, or renders the scene, to a viewer.
Owner:NOKIA TECHNOLOGIES OY

End-to-end learning-based dynamic point cloud coding framework

Some embodiments of a method may include: decoding a motion feature by accessing a motion bitstream; predicting a predicted feature based on the motion feature and one or more reference point cloud frames; decoding a first feature representing an occupancy status of a child level voxel; predicting a second feature based on the first feature and the predicted feature; and decoding a tree voxel occupancy status of the child level voxel via the second feature.
Owner:INTERDIGITAL VC HOLDINGS INC

Generative video coding and decoding method based on multi-modal large model

The invention relates to the technical field of video coding and decoding, and discloses a multi-mode large model-based generative video coding and decoding, which comprises a key frame selection module for determining a key frame by analyzing the semantic and motion characteristics of a video frame; the multi-modal semantic description generation module is used for generating semantic description according to the key frame and the video clip; the key frame compression module is used for realizing efficient compression through latent variable modeling and entropy coding; the key frame reconstruction module reconstructs a key frame by using a conditional diffusion model in combination with the compressed data and the semantic description information; and the video generation module generates a non-key frame by using the semantic description and the key frame, and reconstructs a complete video. Through key frame screening combining semantic and motion information, key frame compression and reconstruction based on a conditional latent variable diffusion model, and frame supplementation and frame insertion generation based on semantic description, efficient compression and high-quality video reconstruction can be realized under a low code rate, and the video storage efficiency and the visual quality are effectively improved.
Owner:上海芯开技术有限公司

Medical image segmentation method and system based on image-text interaction

The invention discloses a medical image segmentation method and system based on image-text interaction, and relates to the technical field of image segmentation, and the method comprises the steps: firstly obtaining a medical image, extracting a multi-scale visual feature, carrying out the deep analysis of a user text description through a medical knowledge graph, and carrying out the knowledge enhancement through a medical anatomical knowledge graph; and thus, an enhanced text vector fusing deep semantics and precise anatomical context is constructed. Furthermore, through a multi-granularity semantic grounding and collaborative fusion mechanism, progressive cross-modal alignment and information interaction are carried out on enhanced text vectors and multi-scale visual features, model focusing is guided, a target area is accurately positioned, and finally a high-precision segmentation mask is generated by a decoder. Therefore, the flexibility of the natural language and the accuracy of the medical priori knowledge are combined, the segmentation challenge in a complex or fuzzy scene can be effectively overcome, and the accuracy and robustness of the segmentation task are remarkably improved.
Owner:ZHEJIANG FEITU IMAGING TECH CO LTD

Depth joint source channel coding method and related device

The deep joint source channel coding method comprises the following steps: a sending end obtains a first semantic feature of a first image, and performs constellation mapping modulation on the first semantic feature to obtain a first initial constellation point; the sending end extracts semantic information of the first initial constellation point, generates a rotation angle corresponding to the semantic information, and rotates the first initial constellation point according to the rotation angle to obtain a first rotation constellation point; the sending end generates a real part interleaving matrix and an imaginary part interleaving matrix based on the first rotating constellation point and the real part and the imaginary part of the channel state information, resorts the real part and the imaginary part of the symbol sequence respectively and then recombines the real part and the imaginary part into a complex signal to obtain a first interleaving signal; the sending end transmits the first interleaved signal to a receiving end through a fading channel, and the receiving end receives a second interleaved signal; the receiving end performs semantic reconstruction decoding on the second interleaved signal to obtain a second image; or, the receiving end performs semantic classification decoding on the second interlaced signal to obtain a second image category.
Owner:BEIJING UNIV OF POSTS & TELECOMM

Anti-compression coding robust video watermark generation method based on adversarial neural network

The invention is suitable for the field of digital watermarking, and provides an anti-compression coding robust video watermark generation method based on an adversarial neural network, and the method comprises the steps: constructing an MSCA-GAN model; the model comprises an encoder, a decoder, a distortion layer, a discriminator network and an opponent network, and the specific steps are as follows: step S1, the encoder extracts different scale features of a video frame by using a multi-scale convolution attention mechanism, calculates attention weights, efficiently hides and embeds binary watermark information into the video, generates a watermark-containing video, and transmits the watermark-containing video to the decoder; performing confrontation optimization on an embedding strategy with a discriminator in training; according to the method, for H.264 compression layer special training, the multi-scale convolution attention mechanism and the depth separable convolution are combined, and the anti-compression robustness of the watermark under the H.264 standard is effectively improved; through common attack training such as noise layer simulation cutting and zooming, the watermark can still keep high extraction accuracy and robustness in a complex environment.
Owner:ENG UNIV OF THE CHINESE PEOPLES ARMED POLICE FORCE

PCB circuit microdefect real-time detection algorithm optimization method based on artificial intelligence

The invention discloses a PCB circuit microdefect real-time detection algorithm optimization method based on artificial intelligence, and relates to the field of printed circuit board monitoring, and the method comprises the steps: collecting multi-physical field data including a temperature field, a stress field, current density distribution and corrosion medium concentration in real time; generating time-space synchronous multi-modal fusion data; quantifying a thermal-mechanical-electric-chemical field interaction relationship through an inter-field coupling coefficient matrix; generating a four-dimensional synergistic effect matrix through tensor fusion; and based on the four-dimensional synergistic effect matrix, in combination with a spatio-temporal evolution diagram output by the carbonization path prediction model, establishing a three-dimensional mapping relationship among the crack size, the carbonization path and the material residual strength, and predicting a residual life prediction value. The method has the advantages that through real-time synchronous fusion of multi-physical field data and tensor coupling modeling, accurate tracking of a crack propagation path and dynamic evaluation of the residual life are achieved, the detection sensitivity is improved, and the life prediction error is reduced.
Owner:DICKSON CIRCUITS (SHENZHEN) LTD

Pharmaceutical hyperspectral reconstruction method based on coded aperture snapshot spectral imaging system

A pharmaceutical hyperspectral reconstruction method based on a coded aperture snapshot spectral imaging (CASSI) system includes: collecting and processing original pharmaceutical hyperspectral images to obtain augmented pharmaceutical hyperspectral images; performing simulated spatial encoding on the augmented pharmaceutical hyperspectral images to obtain encoded measurement images; performing spectral inverse shift on the encoded measurement images, then performing inverse encoding to obtain inversely encoded three-dimensional hyperspectral images, using the augmented pharmaceutical hyperspectral images as target images, and constructing a training set and a testing set according to the inversely encoded three-dimensional hyperspectral images and the target images; constructing a deep symmetric neural reconstruction network, and training and testing the deep symmetric neural reconstruction network; and deploying a tested deep symmetric neural reconstruction network onto the CASSI system, real-time collecting pharmaceutical measurement images using the snapshot coded imaging system, and performing computational reconstruction on the pharmaceutical measurement images to obtain reconstructed three-dimensional hyperspectral images.
Owner:HUNAN UNIV

Video snapshot compression imaging reconstruction method based on space-time deformable attention

The invention provides a video snapshot compression imaging reconstruction method based on spatio-temporal deformable attention, which improves the reconstruction quality and efficiency, and comprises the following steps: inputting a single frame compression measurement value and a measurement matrix into an initial reconstruction module to obtain an initial reconstruction video frame; inputting the initial reconstructed video frame into a feature extraction encoder, mapping the initial reconstructed video frame to a high-dimensional feature space through multi-layer 3D convolution, and outputting a feature map; the feature map is input into a plurality of stacked DenseRNet Blocks, and the number of the DenseRNet Blocks is one; the DenseRNet Block internally comprises a plurality of DeT Blocks, after the DenseRNet Block dynamically divides an input feature channel, grouping progressive processing and feature fusion are carried out through the plurality of DeT Blocks, and the DeT Blocks comprise a deformable space convolution branch used for modeling local deformation perception, a time self-attention branch used for modeling global time sequence dependence and a feature interaction module used for cross-channel information interaction; and the features processed by the DenseRNet Block are input into a video reconstruction decoder, and a reconstructed video sequence is output through up-sampling of transposition convolution and refining of multilayer 3D convolution.
Owner:DALIAN UNIV

Limited angle CT reconstruction method based on combination of three-dimensional conditional diffusion model and synchronous iteration

The invention belongs to the field of CT (Computed Tomography) tomography reconstruction technology and artificial intelligence, and discloses a finite angle CT reconstruction method based on combination of a three-dimensional conditional diffusion model and synchronous iteration. The CT is an imaging technology which utilizes X-rays to irradiate a target from different angles and acquire projection, and obtains an internal three-dimensional structure through reconstruction. Different from traditional CT depending on nearly full-angle scanning, the method only collects limited-angle projection, and achieves fault reconstruction under the limited-angle condition through a three-dimensional condition diffusion model of space domain-frequency domain two-way decoding and by means of structural generality priori of a workpiece. Meanwhile, projection and fault data consistency correction is carried out in combination with a synchronous iteration reconstruction technology, the advantages of an iteration method in the aspect of physical mechanism characterization are exerted, interpretable physical constraints are provided for a deep learning network, and therefore the reliability and precision of a reconstruction result are improved.
Owner:DALIAN UNIV OF TECH

Point cloud data processing method and apparatus

In a method for processing point cloud data according to embodiments, point cloud data can be encoded and transmitted to a bitstream. In a method for processing point cloud data according to embodiments, a bitstream comprising point cloud data can be received, and the point cloud data can be decoded.
Owner:LG ELECTRONICS INC