Comic Image Generation Using Pose Assistance Images
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional methods for converting text content into comics suffer from inaccuracies, leading to compromised generation effects in comic images.
Innovation Solution
A comic image generation method that acquires pose text information to describe target action poses, determines a pose assistance image with matching reference action poses, and uses an artificial intelligence model to generate comic images, thereby enhancing the accuracy of action poses and comic images.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Manufacturing precision
If conventional text-to-comic conversion methods are used, then the process is simple, but the accuracy of action poses and comic images deteriorates
Solution Approach 1:
The patent segments the comic image generation process into distinct components: text information acquisition, pose text information extraction, pose assistance image determination, and AI-based comic image generation. This segmentation allows each component to be optimized independently, improving pose accuracy while managing overall system complexity.
Solution Approach 2:
The patent performs preliminary actions by extracting pose text information from comic text information before generating the final comic images. This preliminary extraction of action pose descriptions enables the AI model to focus on accurate pose generation, thereby improving manufacturing precision in the main generation process.
2Manufacturing precision
If pose text information and pose assistance images are combined, then the accuracy of comic images improves, but the complexity of the generation process increases
Solution Approach 1:
The patent merges pose text information and pose assistance images as combined inputs to the AI model. This merging of multiple information sources provides comprehensive guidance for the AI model, significantly improving the accuracy of generated comic images while the modular architecture manages the increased process complexity.
Solution Approach 2:
The patent introduces pose text information as an intermediary that bridges the gap between comic text information and the AI model. This intermediary translates textual descriptions into structured pose guidance, enabling more accurate image generation without directly increasing the complexity of the AI model itself.
3Reliability
If pose assistance images with matching reference action poses are determined, then the reliability of pose generation improves, but the time required for image search and processing increases
Solution Approach 1:
The patent performs preliminary determination of pose assistance images based on extracted pose text information before the main comic image generation. This preliminary action ensures that reliable reference poses are prepared in advance, improving the reliability of pose generation while allowing parallel processing to minimize time loss.
Solution Approach 2:
The patent implements a feedback mechanism where pose text information is extracted from comic text information, used to determine pose assistance images, and then fed back into the AI model for refined comic image generation. This feedback loop continuously improves pose generation reliability while the automated feedback process minimizes manual intervention time.
Data Source
AI summary
The present disclosure provides a comic image generation method and apparatus, a computer device, and a storage medium, and the method includes: acquiring pose text information corresponding to a comic storyboard, the pose text information being used for describing target action poses of at least one target object; determining a pose assistance image corresponding to the comic storyboard according to the pose text information, reference action poses of a reference object in the pose assistance image matching the target action poses; and generating comic images corresponding to the comic storyboard using an artificial intelligence model according to the pose text information and the pose assistance image, target objects in the comic images having the target action poses.


