Text-Guided 3D Model View Generation with Contour Control

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing three-dimensional model generation techniques face challenges in controllability and interactive modification, particularly when using single-image or text-driven methods, leading to poor controllability and difficulty in making quick, partial modifications.

Innovation Solution

A method and apparatus that utilize a three-dimensional geometric model and text description to generate views of a target three-dimensional model with texture, allowing for a certain degree of freedom in contour similarity and enabling interactive editing and quick preview of modifications through 3D and 2D control mechanisms.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If single-image or text-driven methods are used for three-dimensional model generation, then generation speed is improved, but controllability deteriorates

Engineering Contradiction:
Improvegeneration speedVSAvoidcontrollability
Core Design Contradiction:
ProductivityVSEase of operation

Solution Approach 1:

The patent segments the control process into two independent parts: a three-dimensional geometric model for spatial structure control and a two-dimensional image for surface appearance control. This segmentation allows each component to be optimized independently, maintaining generation speed while improving controllability through precise 3D spatial control.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces a neural radiance field as an intermediary between the geometric model and the final rendered image. This intermediary enables efficient rendering and allows for interactive modification by mediating between the 3D structure and 2D appearance requirements, thus improving controllability without sacrificing generation speed.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Productivity

If existing three-dimensional model generation techniques are used, then model generation is achieved, but ease of modification deteriorates

Engineering Contradiction:
Improvemodel generation capabilityVSAvoidease of modification
Core Design Contradiction:
ProductivityVSEase of operation

Solution Approach 1:

The patent implements dynamic interactivity by allowing real-time modification of the three-dimensional geometric model and corresponding updates to the rendered views. The system supports interactive adjustment of model parameters, camera poses, and material properties, enabling users to quickly modify and iterate on design changes.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent enables easy modification by allowing users to adjust key parameters such as camera poses, geometric model properties, and material characteristics. These parameter changes are processed through the neural radiance field to generate updated views efficiently, making the modification process simple and rapid.

Inventive Principle:
Principle #35Parameter changes

3Adaptability or versatility

If text description is used for model generation, then model diversity is improved, but manufacturing precision deteriorates

Engineering Contradiction:
Improvemodel diversityVSAvoidcontour accuracy
Core Design Contradiction:
Adaptability or versatilityVSManufacturing precision

Solution Approach 1:

The patent merges the advantages of text-driven generation (diversity) with 3D geometric modeling (precision) by combining a three-dimensional geometric model with a text description. The geometric model provides precise contour control while the text description enables diverse material and appearance variations, achieving both precision and diversity simultaneously.

Inventive Principle:
Principle #5Merging (Combining)

Data Source

PatentUS20250299448A1Method and apparatus for generating views of three-dimensional model, electronic device, and storage medium
Publication Date: 2025.09.25 BEIJING ZITIAO NETWORK TECH CO LTD
  • US20250299448A1 patent drawing
  • US20250299448A1 patent drawing
  • US20250299448A1 patent drawing

AI summary

The present disclosure provides a method and an apparatus for generating views of a three-dimensional model, an electronic device, and a storage medium. The method for generating views of a three-dimensional model includes: obtaining a three-dimensional geometric model and a text description; and generating views of a target three-dimensional model based on the geometric model and the text description, wherein the target three-dimensional model has texture information and the target three-dimensional model conforms to the text description, a similarity between a contour of the target three-dimensional model and a contour of the geometric model is greater than a preset similarity, and the views of the target three-dimensional model include: views corresponding to first camera poses, the number of the first camera poses being one or more.