Image Generation Support Using Style and Element Image Inputs
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing image generation technologies, such as those described in Patent Literature 1, are limited to generating images of characters in arbitrary postures and do not support the generation of various images.
Innovation Solution
A generation support device and method that includes a generation information acquisition unit to acquire style and element image information from a user, and a result image generation unit to input this information into a generative model to generate the desired image.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If existing image generation technology (Patent Literature 1) is used, then character images in arbitrary postures can be generated, but the technology is limited and not applicable to generation of various images
Solution Approach 1:
The patent creates a universal image generation system that can handle multiple types of images (characters, products, scenes, etc.) through a single integrated platform. The system accepts various input types including text descriptions, reference images, and style examples, and can generate diverse output types such as character images, product images, and scene images, making it applicable to wide ranges of generation tasks without requiring separate specialized systems for each image type
2Adaptability or versatility
If a generative model is used to generate diverse images, then various types of images can be created, but the process becomes more complex requiring multiple input parameters
Solution Approach 1:
The patent introduces an intermediary processing layer between user input and the generative model that automatically synthesizes multiple required parameters. The system accepts simple user inputs (text descriptions or reference images) and automatically generates the necessary style embeddings, content representations, and other parameters needed by the generative model, shielding users from the complexity of parameter specification while enabling diverse image generation
Solution Approach 2:
The system performs preliminary processing of input data by pre-computing style embeddings and content representations before they are needed by the generative model. By preparing these intermediate representations in advance through automatic feature extraction and encoding, the system simplifies the user interaction process while maintaining the capability to generate diverse image types
Data Source
AI summary
Provided is a generation support device for supporting generation of a result image. The generation support device includes: a generation information acquisition unit configured to acquire, from a user, generation information including at least style information regarding a style of the result image and an element image constituting part of the result image; and a result image generation unit configured to input text information generated based on the generation information into a generative model, and generate the result image based on output information output from the generative model.


