Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

47 results about "Multiple viewpoint" patented technology

Large-size measurement-oriented anti-shielding three-dimensional scanning measurement field construction method

The invention discloses an anti-shielding three-dimensional scanning measurement field construction method for large-size measurement, and belongs to the technical field of optical three-dimensional measurement. The method comprises the following steps: firstly, performing adaptive sampling based on geometric features of a workpiece to generate a candidate mark point set; secondly, constructing a virtual measurement scene, simulating a dynamic scanning process by using a ray tracing technology, and calculating the visibility and observation quality of each point under multiple viewpoints; then measuring precision is quantified by using a Fisher information matrix, and a layout optimization model which takes maximization of information gain as a target and meets coverage rate and engineering constraint at the same time is established; and finally, solving the model to obtain an optimal mark point subset, and verifying and outputting the optimal mark point subset. According to the method, the problems of measurement interruption, low point distribution efficiency and difficulty in guaranteeing precision caused by shielding of a large-size complex component in scanning are solved, and automatic construction of a high-robustness and high-precision measurement field is realized.
Owner:NORTHWESTERN POLYTECHNICAL UNIV

Comment generation method and device, equipment, storage medium and program product

PendingCN121766420ASemantic analysisMachine learningEngineeringMultiple viewpoint
The invention relates to a comment generation method and device, equipment, a storage medium and a program product. According to one embodiment of the disclosure, the method comprises: selecting a target personality portrait from a pre-constructed personality portrait database, the personality portrait database comprising a plurality of different personality portraits; selecting a target watching point matched with the target personality portrait from watching points of multiple angles corresponding to original contents, wherein the original contents comprise contents for generating comments; and inputting the target personality portrait and the target watching point into a pre-trained comment generation model to obtain a currently generated comment. According to the method and the device, the diversity of the generated comments can be improved, so that the network environment interaction atmosphere foiling effect can be improved.
Owner:BEIJING XIAOMI MOBILE SOFTWARE CO LTD

A target recognition processing method based on a sliding window and multi-view fusion

This invention relates to the field of visual recognition technology, specifically disclosing a target recognition processing method based on sliding window and multi-view fusion. The method includes: acquiring video streams from multiple viewpoints of a target scene; extracting and recognizing visual features of video frames for each viewpoint video segment to obtain basic recognition results under a single viewpoint; determining whether the basic recognition results of multiple consecutive frames under each single viewpoint conform to preset business rules, generating a single-view result sequence; performing frame-level fusion of the visual features under each single viewpoint to obtain fused visual features of multiple frames under multiple viewpoints; determining whether the fused visual features of each frame conform to preset spatial rules, generating a multi-view result sequence; and performing smoothing and fusion processing on each single-view result sequence and the multi-view result sequence to generate the recognition result of the current video segment. This invention can fully utilize multi-view data to improve recognition accuracy.
Owner:DMAI (GUANGZHOU) CO LTD

Exposure sequence adaptive optimization method for multi-view automatic measurement

The invention relates to an exposure sequence adaptive optimization method and device for multi-view automatic measurement, computer equipment, a computer readable storage medium and a computer program product. The method comprises the following steps: acquiring an original exposure time sequence of a plurality of viewpoints; the original exposure time sequence comprises a plurality of exposure segments, and each exposure segment corresponds to one viewpoint; calculating similarity data among the exposure segments, and grouping the exposure segments according to the similarity data to obtain a plurality of initial exposure groups; for each initial exposure group, re-grouping the exposure segments at the boundary position to obtain a plurality of boundary grouping results, and calculating grouping cost data for each boundary grouping result; and selecting a target grouping result from the boundary grouping results according to the grouping cost data, optimizing the target grouping result according to the grouping cost data to obtain a plurality of target exposure groups, and calculating the exposure time of each target exposure group. The method can shorten the measurement period.
Owner:WUHAN POWER3D TECH

Image processing device, image processing method, and program

We estimate a three-dimensional field capable of generating virtual viewpoint images that do not exhibit any noticeable inconsistencies in image quality when switching virtual viewpoints. [Solution] The image processing device 102 acquires multiple captured images obtained from multiple viewpoints and camera parameters corresponding to each viewpoint, acquires the shooting resolution which is the resolution of each captured image, generates multiple training images corresponding to each captured image by converting the resolution of each captured image based on the resolution of each of the multiple training images corresponding to each captured image which is determined based on the shooting resolution of each captured image, generates camera parameters corresponding to each training image by converting the camera parameters corresponding to each viewpoint based on the resolution of each training image, and learns a learning model concerning the three-dimensional field of the target space based on the multiple training images and the camera parameters corresponding to each training image.
Owner:CANON KK

Video, image and 3D information generation method and device, medium and product

The embodiment of the invention provides a video, image and 3D information generation method and device, a medium and a product, and belongs to the field of AI.The method comprises the steps that multiple to-be-edited images corresponding to multiple view angles are obtained, a to-be-edited area in a first to-be-edited image of an initial view angle is edited to obtain a first edited image, and the first edited image is edited to obtain a second edited image; and generating a plurality of sheltered second to-be-edited images corresponding to the other plurality of second to-be-edited images according to the to-be-edited area in the first to-be-edited image. And inputting the first edited image and the plurality of occluded second to-be-edited images into a target image restoration model to obtain a plurality of second edited images, wherein the target image restoration model diffuses an editing result in the first edited image to the plurality of occluded second to-be-edited images. And generating a video according to the plurality of edited images. Through the scheme, multi-view consistent image editing can be realized.
Owner:ALIBABA (CHINA) CO LTD

Dynamic digital human high-fidelity real-time rendering method and device, equipment and storage medium

This application relates to a high-fidelity real-time rendering method, apparatus, device, and storage medium for dynamic digital humans. It includes receiving multimodal driving signals, mapping them to expression and pose parameters, inputting a parameterized human model to generate standard spatial geometry and skinning weights, and constructing a three-dimensional Gaussian sputtering field to deform to the target pose to obtain dynamic geometric data; arranging sparse virtual cameras in the target pose space, and simultaneously rendering and generating a sparse reference view set containing color and depth within a single rendering call; determining dense virtual viewpoints based on the display device's viewpoint parameters, and generating depth range textures for each dense viewpoint through downsampling depth reprojection and point sputtering techniques; determining ray step intervals using the depth range textures, sampling along the rays and fusing signed distance values ​​truncated from multiple viewpoints to generate surface position textures; projecting the surface positions onto sparse view sampled colors, fusing to generate a dense view color image, and adapting it for output to different display devices.
Owner:GUANGZHOU WANQU MEDIA TECHNOLOGY CO LTD

A robust multi-view skeleton fusion method based on information prioritization and effective joint masking

This invention belongs to the field of computer vision and action recognition, specifically relating to a robust multi-view skeleton fusion method based on information priority selection and effective joint masking. It includes: acquiring skeleton data from multiple viewpoints, performing data preprocessing and effective joint extraction; skeleton registration based on information content to obtain a registered skeleton; obtaining the skeleton most suitable for action recognition based on the registered skeleton and a GCN model; obtaining the action category based on the most suitable skeleton using a modern action recognition model; and adjusting model parameters using cross-entropy loss based on the action category results to optimize iterative recognition accuracy. This invention effectively solves the problem of low-quality noise interfering with high-fidelity data in multi-view fusion through an information content evaluation mechanism and a Boolean mask completion strategy; combined with GCN adaptive viewpoint adjustment, it enhances the model's robustness to different camera layouts and significantly improves the accuracy and stability of human action recognition in complex occlusion environments.
Owner:NANTONG UNIV

A three-dimensional indoor scene generation method and related device

The application discloses a three-dimensional indoor scene generation method and related equipment, and the method comprises the steps of obtaining a text description of an indoor scene; generating an indoor layout through a preset diffusion model; converting the indoor layout into a two-dimensional height field and a two-dimensional semantic graph; predicting a neural radiance field based on the two-dimensional height field and the two-dimensional semantic graph; sampling multiple viewpoints in the predicted neural radiance field, rendering the multiple viewpoints, and obtaining an RGB-D image corresponding to each viewpoint, wherein RGB is color information of a two-dimensional image, and D is depth information of the two-dimensional image; constructing a truncated signed distance field body according to the RGB-D image corresponding to each viewpoint, generating a grid triangular facet based on the truncated signed distance field body, and obtaining an indoor three-dimensional scene. The application can generate a high-quality three-dimensional indoor scene without real multi-view images, improve the efficiency of generating a three-dimensional scene, and can be widely applied in the technical field of computer vision.
Owner:SUN YAT SEN UNIV

Volumetric performance capture with neural rendering

Example embodiments relate to techniques for volumetric performance capture with neural rendering. A technique may involve initially obtaining images that depict a subject from multiple viewpoints and under various lighting conditions using a light stage and depth data corresponding to the subject using infrared cameras. A neural network may extract features of the subject from the images based on the depth data and map the features into a texture space (e.g., the UV texture space). A neural renderer can be used to generate an output image depicting the subject from a target view such that illumination of the subject in the output image aligns with the target view. The neural render may resample the features of the subject from the texture space to an image space to generate the output image.
Owner:GOOGLE LLC

Holographic communication method, electronic device, storage medium, and program product

PendingCN122289552APattern recognitionMultiple viewpoint
This document discloses a holographic communication method, electronic device, storage medium, and program product. The method includes: acquiring a three-dimensional model of a first object; the three-dimensional model is obtained based on multiple depth images and multiple attribute images, the multiple depth images and multiple attribute images being obtained based on multiple first images, the multiple first images being acquired from multiple viewpoints of the first object, the multiple attribute images including attribute data of multiple primitives under the multiple viewpoints, the primitives being basic units representing the first object; and rendering and displaying the three-dimensional model based on the three-dimensional model and the client's viewpoint. Therefore, high-fidelity stereoscopic image reconstruction can be achieved on the client using the acquired multi-view images, while ensuring viewpoint integrity.
Owner:BEIJING ZITIAO NETWORK TECH CO LTD

System and method for collaboration when viewing aerial imagery and derived content

A system and method for viewing aerial imagery and derived content on a web-based imagery browser that can be accessed by at least two users simultaneously. In addition, a system for interactively generating 3D geometry data for a location imaged from multiple viewpoints.
Owner:NEARMAP US INC

System

A system is provided.SOLUTION: A system, comprising: means for receiving a request from a user; means for retrieving and generating information based on the request; and means for providing the generated information to the user, wherein the means for receiving the request from the user includes means for receiving the request from the user, means for generating the information based on the request, and means for providing the information to the user. AI AI AI.SELECTED DRAWING: Figure 1
Owner:SOFTBANK GROUP CORP

Apparatus and method for generating object-based stereoscopic images

According to an embodiment of the present disclosure, a method of generating a stereoscopic video includes: obtaining video data; extracting at least one multiple viewpoint video data from the video data; identifying an object associated with the video data; generating tracking information for the object based on the at least one multiple viewpoint video data; and performing rendering on the object based on a depth value corresponding to the tracking information.
Owner:ELECTRONICS & TELECOMM RES INST

Medical assistant robot and adaptive scan view planning method and system thereof

PendingCN122473401ASimulationMesh grid
The application provides a kind of adaptive scanning visual angle planning method of medical auxiliary robot, mainly includes four steps: first, the three-dimensional grid model of patient is normalized to simulation environment, and the operating space of camera is limited;Second, according to three-dimensional grid model, target site identification and historical scanning information, construct composite state vector;Then, input the vector through strategy network, generate three-dimensional action vector representing the next optimal viewpoint coordinates;Finally, generate a series of viewpoints composed of multiple viewpoints, based on this sequence, use path planning algorithm to generate collision-free trajectory, and collect data at each viewpoint, finally obtain the three-dimensional point cloud of target site.The application has the beneficial effects: through conditional learning, a model with strong generalization ability for multiple parts is realized, and multiple targets such as scanning integrity, quality and safety can be automatically weighed, and finally a safe, efficient and high-precision automatic scanning scheme far beyond traditional methods is generated.
Owner:TIANJIN UNIVERSITY OF TECHNOLOGY +2

Virtual viewpoint image synthesis device and method using quasi-uniform spherical coordinate grid

An image synthesis device includes a memory storing a virtual viewpoint image synthesis program, and a processor configured to execute the program. The program receives images of multiple viewpoints obtained by filming around a user, position information on the user, and information on a direction in which scene is viewed, estimates color and density by inputting the position information of the user and the direction in which the object is viewed to a neural radiance field model constructed by using the images of the multiple viewpoints, and synthesizes images of a virtual viewpoint by performing volume rendering by using estimated color and density, the neural network radiance field model is constructed based on a quasi-uniform spherical coordinate system, and the quasi-uniform spherical coordinate system includes a sum of a Yin grid and a Yang grid and has a grid structure that increases exponentially in a radial direction.
Owner:SEOUL NATIONAL UNIVERSITY R&DB FOUNDATION

Collaborative scanning method, device, medium and product for an aeroengine

PendingCN122312927AAviationPoint cloud
This invention discloses a collaborative scanning method, device, medium, and product for aero-engines. The method includes: controlling a continuous inspection robotic arm to perform a coarse scan of the aero-engine to obtain a coarse point cloud model; forming multiple candidate viewpoints based on the coarse point cloud model; grouping the candidate viewpoints to obtain multiple viewpoint groups and forming a joint viewpoint visibility matrix; obtaining a sequence of target viewpoint groups based on the joint viewpoint visibility matrix; controlling each industrial camera of the robotic arm to position itself at a target viewpoint in each target viewpoint group for high-resolution imaging by adjusting the bending and / or rotating traction ropes of the continuous inspection robotic arm, thereby obtaining a local fine point cloud; and obtaining a high-precision 3D solid model based on the coarse point cloud model and the local fine point cloud. This invention fundamentally avoids mechanical interference and collision damage, achieves full coverage of complex curved surfaces without blind spots, and improves the accuracy and efficiency of 3D reconstruction of aero-engines.
Owner:CIVIL AVIATION UNIV OF CHINA

Multi-view 3D Reconstruction Method with Fusing Edge Priors

PendingCN122312872AVisual technologyEdge maps
This invention relates to a multi-view 3D reconstruction method incorporating edge priors, belonging to the field of computer vision technology, and solves the problem of poor accuracy in existing 3D reconstruction results. The 3D reconstruction method includes: acquiring multiple viewpoint images of the scene to be reconstructed; selecting one viewpoint image and its corresponding edge image from the multiple viewpoint images as a reference image and its edge image, and selecting multiple source images from the multiple viewpoint images to obtain multiple image combinations; inputting each of the multiple image combinations into a pre-trained depth prediction model to obtain a depth map corresponding to the reference image in each image combination; and converting the pixels in the reference image into spatial points in the scene to be reconstructed based on the intrinsic and extrinsic parameter matrices of the reference image and the corresponding depth map in each image combination, thereby obtaining the reconstructed scene. This improves the accuracy of 3D reconstruction.
Owner:INST OF MICROELECTRONICS CHINESE ACAD OF SCI LTD

Display device

This invention provides a technology that reduces user discomfort with light emanating from pixels in a display device capable of displaying different images from multiple viewpoints. [Solution] The display device includes a quasi-planar substrate including a flat or curved surface; a group of pixels including a plurality of display pixels at different positions in a first direction parallel to the tangent plane of the substrate; a first lens element configured to refract light incident from a plurality of the display pixels of the pixel group in a third direction perpendicular to the tangent plane of the substrate in separate directions within a plane including the first and third directions; and a second lens element configured to refract light incident from the display pixels in the third direction within a plane including a second direction and a third direction that intersect the first direction at a specific angle within the tangent plane of the substrate.
Owner:DUAL MOVE CO LTD

Multi-view cooperative job behavior risk assessment method and system thereof

ActiveCN121999539BHuman bodySpatial mapping
The application discloses a multi-viewpoint cooperative work behavior risk assessment method and system, relates to the technical field of behavior risk assessment, and comprises the following steps: synchronously acquiring an initial image sequence of multiple viewpoints around a work station, extracting human body key point 2D skeleton feature data of each viewpoint, and calling environment semantic map data containing device entity 3D geometric boundary and device operation point coordinates; each 2D skeleton feature data is projected into a world coordinate system consistent with the environment semantic map through spatial coordinate mapping transformation, the spatial mapping coordinates of human body key points under each viewpoint are obtained, and the spatial mapping coordinates of different viewpoints belonging to the same human body target are spatially aggregated to generate an initial 3D fusion skeleton flow; the beneficial effects are that non-real action data generated due to shielding or perspective deviation can be accurately identified and removed, and the false alarm problem caused by the absence of logical verification in a complex space in a traditional algorithm is solved.
Owner:XIAN UNIV OF TECH

Multi-dimensional demonstration method and device based on time domain technology and storage medium

ActiveCN122152779BTime domainAlgorithm
This application discloses a multi-dimensional demonstration method, apparatus, and storage medium based on temporal domain technology, relating to the field of data interaction technology. This application first acquires a temporal file containing multiple viewpoints and / or multiple time points, and encapsulates multi-dimensional index information including spatial and temporal indices. Then, based on the multi-dimensional index information, a demonstration narrative stream is constructed, including a sequence of keyframes for the demonstration and the camera motion paths of the keyframes. During playback rendering, the current camera state is acquired in real time, and combined with the predicted information of the camera motion paths, the target display area within the current view frustum is calculated. Finally, the corresponding target pyramid tile data is retrieved from the multi-dimensional index information and progressively decoded and rendered in order of increasing resolution, achieving seamless loading of the demonstration scene from an overall overview to local micro-details. This improves the loading efficiency of the multi-dimensional demonstration scene and reduces resource consumption.
Owner:SI CHUAN ZHONG SHENG MATRIX TECH DEV CO LTD +1

Camera calibration method and calibration system

This eliminates the need to acquire pattern images from multiple viewpoints to obtain the camera's optical parameters, improving robustness against shifts in shooting position. [Solution] A camera calibration method that acquires a pattern image obtained by photographing calibration boards 30 and 32, each equipped with two spaced-apart pattern boards, using cameras 10, 12, and 14; obtains principal point coordinates indicating the position of the lens optical axis on the image sensor based on the patterns displayed on the two pattern boards captured in the pattern image; and uses the principal point coordinates and the relative positional relationship between cameras 10, 12, and 14 and the calibration boards 30 and 32 to acquire optical parameters other than the principal point coordinates, including the focal length, based on the pattern image.
Owner:ASTEMO LTD

Building image clustering device and building footprint generation device

From a set of images of numerous buildings, each taken from multiple viewpoints, the image data of the same building is appropriately clustered into a single cluster. [Solution] The building image clustering device 1 receives an image data receiving unit 105 that takes images of two or more buildings from multiple locations, a feature vector acquisition unit 106 that acquires a feature vector for each image data by using a visual language model on the image data group, and a rastering unit 107 clusters the image data group into clusters based on the feature vectors, with groups of image data that are estimated to have been taken of the same building being treated as one cluster.
Owner:MICWARE CO LTD

Three-dimensional super resolution using generative video models

In implementing three-dimensional (3D) super-resolution techniques using generative video models, a processing device receives a first 3D representation of an object. The processing device generates an intermediate video of the object from multiple viewpoints of the first 3D representation. A machine-learning model then generates an upsampled video from the intermediate video. The upsampled video is in a higher resolution than the intermediate video. In one example, the machine learning model is a video-based generative upsampler. Based on 3D reconstruction of the object from the upsampled video, the processing device outputs a second 3D representation of the object in a resolution higher than the first 3D representation.
Owner:ADOBE INC

A method for dynamic inspection of a fan and an electronic device

This invention provides a method and electronic device for dynamic inspection of wind turbines. The method includes: determining a target viewpoint on the radius of the wind turbine rotor, wherein there is a redundant distance between the target viewpoint and the blade tip position of the wind turbine rotor radius, and the wind turbine rotor radius is the horizontal radius of the circular wheel generated by the rotation of the wind turbine blades; planning multiple viewpoints between the center of the wind turbine rotor radius and the target viewpoint based on the target viewpoint; translating the multiple viewpoints according to the real-time yaw direction and anti-yaw direction of the wind turbine to obtain inspection waypoints on the front and back of the wind turbine; and controlling a rangefinder to measure the distance to the wind turbine blades when a drone hovers over each inspection waypoint, and controlling the drone camera to take pictures based on the ranging results. This solves the problem of low efficiency in planning inspection waypoints for dynamic wind turbine inspections in the prior art.
Owner:BEIJING DEEPERCEPTION TECHNOLOGY CO LTD

Triangular mesh-based micro-renderable animatable face three-dimensional reconstruction method

The invention discloses an animatable face three-dimensional reconstruction method capable of micro-rendering based on a triangular mesh, and the method comprises the steps: carrying out the single-frame reconstruction and sequence frame reconstruction successively by taking a character head portrait video shot at a single view angle or a plurality of view angles as an input; and finally outputting a character head portrait basic model in a triangular grid format, main joint position information and control weight of the human face, a human face expression mixed shape and a plurality of human face coloring chartlets. The face three-dimensional model reconstructed by the method is accurate in geometric structure and high in chartlet definition, supports animation editing through joint parameter and expression mixing, and can meet the requirements of the movie and television and game industries for face animation models.
Owner:ZHEJIANG UNIV

Image target tracking method and system based on view angle updating and storage medium

The invention relates to the technical field of image processing, in particular to an image target tracking method and system based on view angle updating and a storage medium, and the method comprises the steps: evaluating the target capturing performance of a plurality of camera view angles based on historical video data; the target capture performance is constructed into a multi-target optimization function, and an initial view angle and a non-dominant camera view angle for target tracking are screened out; performing clustering processing on all camera visual angles to obtain a plurality of visual angle groups; in each view angle group, determining the view angle novelty of the view angle of the non-dominant camera; and based on the combination result of the novelty of the visual angle and the target capture performance, replacing and updating the initial visual angle of target tracking to form the latest visual angle of target tracking. According to the method, the initial vision is updated in real time in combination with the novelty index, an effective mining mechanism for the potential value of the non-dominant view angle is realized, complementary view angle information is found and provided at the critical moment, and the adaptability in a complex scene in the target tracking process is improved.
Owner:ANHUI NORMAL UNIV

Multiple hologram QR code using multiple viewpoints and multiple focal depths

A multiple hologram QR code using multiple viewpoints and multiple focal depths is provided. A method for generating a multiple 3D code, according to an embodiment of the present invention, divides an original 3D code into a plurality of sub-codes having different depths, arranges the divided sub-codes to be displayed in different directions, and generates the multiple 3D code by combining the arranged sub-codes. Therefore, the multiple hologram QR code is implemented such that display directions and focal distances differ for respective portions, and thus it is impossible to easily duplicate the multiple hologram QR code as with a conventional 2D QR code, and a different scanning method in which viewpoint scanning and focus scanning are combined can be provided.
Owner:KOREA ELECTRONICS TECH INST

Display system

This invention provides a display system that can mitigate the reduction in visibility of the main image when displaying images from multiple viewpoints. [Solution] The display system 1 comprises a computer 13 and a display unit 12a. The display unit 12a displays the work machine 4 and / or the surrounding environment (e.g., transport vehicles and work objects, etc.) as the main image MP. The computer 13 displays the work machine 4 and / or the surrounding environment from a different viewpoint than the main image MP as a secondary image SP on the display unit 12a in an area outside the gaze area A1 of the main image MP.
Owner:KOBELCO CONSTR MASCH CO LTD