3D Spatial Position Calculation from 2D Disposition Knowledge
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current image recognition systems often inaccurately recognize positional relationships between objects and require significant time to prepare recognition models, as they neglect disposition features and rely on two-dimensional data that loses three-dimensional positional information.
Innovation Solution
An information processing device and method that acquires and processes disposition knowledge from document information to calculate spatial positional relationships between objects, using an inter-object positional relationship recognition unit, observation position and orientation recognition unit, and spatial positional relationship calculation unit to generate three-dimensional spatial information, independent of the observation point.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If image recognition is used to recognize object types and positional relationships, then object recognition can be achieved, but recognition accuracy deteriorates due to erroneous recognition of positional relationships
Solution Approach 1:
The patent transitions from two-dimensional image data to three-dimensional spatial information by introducing depth prediction and spatial coordinate transformation. The system converts 2D image coordinates into 3D spatial coordinates using predicted depth information and camera parameters, enabling accurate representation of positional relationships in three-dimensional space rather than being constrained to two-dimensional image planes.
2Productivity
If traditional image recognition models are prepared, then object recognition can be performed, but significant time is required to prepare the recognition model
Solution Approach 1:
The system performs preliminary depth prediction and spatial coordinate transformation to pre-calculate three-dimensional positional relationships. By predicting depth information in advance and transforming coordinates before final recognition, the system reduces the computational burden during model execution and accelerates the overall processing workflow.
3Device complexity
If two-dimensional disposition features are used, then processing is simplified, but three-dimensional positional information is lost
Solution Approach 1:
The patent systematically restores three-dimensional information by integrating depth prediction with coordinate transformation. The process involves predicting depth for each object, calculating three-dimensional coordinates using camera parameters (focal length, principal point, rotation, translation), and representing spatial relationships in 3D space, thereby recovering the depth dimension that is inherently lost in two-dimensional images.
Data Source
AI summary
An information processing device includes: a disposition knowledge acquisition unit configured to acquire disposition knowledge information including type information that indicates types of two or more observation targets observed by an observer and positional relationship information between the observation targets; an inter-object positional relationship recognition unit configured to recognize an object type and a positional relationship in the disposition knowledge information; an observation position and orientation recognition unit configured to recognize an observation position and orientation that are a position and an orientation of observation at a time point at which the disposition knowledge information is generated; and a spatial positional relationship calculation unit configured to calculate spatial positional information between two or more of the observation targets based on the positional relationship and the observation position and orientation.


