Comic Image Generation Using Pose Assistance Images

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional methods for converting text content into comics suffer from inaccuracies, leading to compromised generation effects in comic images.

Innovation Solution

A comic image generation method that acquires pose text information to describe target action poses, determines a pose assistance image with matching reference action poses, and uses an artificial intelligence model to generate comic images, thereby enhancing the accuracy of action poses and comic images.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Manufacturing precision

If conventional text-to-comic conversion methods are used, then the process is simple, but the accuracy of action poses and comic images deteriorates

Engineering Contradiction:
Improveaccuracy of action posesVSAvoidcomplexity of generation method
Core Design Contradiction:
Manufacturing precisionVSDevice complexity

Solution Approach 1:

The patent segments the comic image generation process into distinct components: text information acquisition, pose text information extraction, pose assistance image determination, and AI-based comic image generation. This segmentation allows each component to be optimized independently, improving pose accuracy while managing overall system complexity.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent performs preliminary actions by extracting pose text information from comic text information before generating the final comic images. This preliminary extraction of action pose descriptions enables the AI model to focus on accurate pose generation, thereby improving manufacturing precision in the main generation process.

Inventive Principle:
Principle #10Preliminary action

2Manufacturing precision

If pose text information and pose assistance images are combined, then the accuracy of comic images improves, but the complexity of the generation process increases

Engineering Contradiction:
Improveaccuracy of comic imagesVSAvoidcomplexity of generation process
Core Design Contradiction:
Manufacturing precisionVSDevice complexity

Solution Approach 1:

The patent merges pose text information and pose assistance images as combined inputs to the AI model. This merging of multiple information sources provides comprehensive guidance for the AI model, significantly improving the accuracy of generated comic images while the modular architecture manages the increased process complexity.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The patent introduces pose text information as an intermediary that bridges the gap between comic text information and the AI model. This intermediary translates textual descriptions into structured pose guidance, enabling more accurate image generation without directly increasing the complexity of the AI model itself.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Reliability

If pose assistance images with matching reference action poses are determined, then the reliability of pose generation improves, but the time required for image search and processing increases

Engineering Contradiction:
Improvereliability of pose generationVSAvoidtime for image search and processing
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent performs preliminary determination of pose assistance images based on extracted pose text information before the main comic image generation. This preliminary action ensures that reliable reference poses are prepared in advance, improving the reliability of pose generation while allowing parallel processing to minimize time loss.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent implements a feedback mechanism where pose text information is extracted from comic text information, used to determine pose assistance images, and then fed back into the AI model for refined comic image generation. This feedback loop continuously improves pose generation reliability while the automated feedback process minimizes manual intervention time.

Inventive Principle:
Principle #23Feedback

Data Source

PatentUS20250078333A1Comic image generation method, computer device and storage medium
Publication Date: 2025.03.06 BEIJING ZITIAO NETWORK TECH CO LTD
  • US20250078333A1 patent drawing
  • US20250078333A1 patent drawing
  • US20250078333A1 patent drawing

AI summary

The present disclosure provides a comic image generation method and apparatus, a computer device, and a storage medium, and the method includes: acquiring pose text information corresponding to a comic storyboard, the pose text information being used for describing target action poses of at least one target object; determining a pose assistance image corresponding to the comic storyboard according to the pose text information, reference action poses of a reference object in the pose assistance image matching the target action poses; and generating comic images corresponding to the comic storyboard using an artificial intelligence model according to the pose text information and the pose assistance image, target objects in the comic images having the target action poses.