Text-Guided 3D Model View Generation with Contour Control
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing three-dimensional model generation techniques face challenges in controllability and interactive modification, particularly when using single-image or text-driven methods, leading to poor controllability and difficulty in making quick, partial modifications.
Innovation Solution
A method and apparatus that utilize a three-dimensional geometric model and text description to generate views of a target three-dimensional model with texture, allowing for a certain degree of freedom in contour similarity and enabling interactive editing and quick preview of modifications through 3D and 2D control mechanisms.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If single-image or text-driven methods are used for three-dimensional model generation, then generation speed is improved, but controllability deteriorates
Solution Approach 1:
The patent segments the control process into two independent parts: a three-dimensional geometric model for spatial structure control and a two-dimensional image for surface appearance control. This segmentation allows each component to be optimized independently, maintaining generation speed while improving controllability through precise 3D spatial control.
Solution Approach 2:
The patent introduces a neural radiance field as an intermediary between the geometric model and the final rendered image. This intermediary enables efficient rendering and allows for interactive modification by mediating between the 3D structure and 2D appearance requirements, thus improving controllability without sacrificing generation speed.
2Productivity
If existing three-dimensional model generation techniques are used, then model generation is achieved, but ease of modification deteriorates
Solution Approach 1:
The patent implements dynamic interactivity by allowing real-time modification of the three-dimensional geometric model and corresponding updates to the rendered views. The system supports interactive adjustment of model parameters, camera poses, and material properties, enabling users to quickly modify and iterate on design changes.
Solution Approach 2:
The patent enables easy modification by allowing users to adjust key parameters such as camera poses, geometric model properties, and material characteristics. These parameter changes are processed through the neural radiance field to generate updated views efficiently, making the modification process simple and rapid.
3Adaptability or versatility
If text description is used for model generation, then model diversity is improved, but manufacturing precision deteriorates
Solution Approach 1:
The patent merges the advantages of text-driven generation (diversity) with 3D geometric modeling (precision) by combining a three-dimensional geometric model with a text description. The geometric model provides precise contour control while the text description enables diverse material and appearance variations, achieving both precision and diversity simultaneously.
Data Source
AI summary
The present disclosure provides a method and an apparatus for generating views of a three-dimensional model, an electronic device, and a storage medium. The method for generating views of a three-dimensional model includes: obtaining a three-dimensional geometric model and a text description; and generating views of a target three-dimensional model based on the geometric model and the text description, wherein the target three-dimensional model has texture information and the target three-dimensional model conforms to the text description, a similarity between a contour of the target three-dimensional model and a contour of the geometric model is greater than a preset similarity, and the views of the target three-dimensional model include: views corresponding to first camera poses, the number of the first camera poses being one or more.


