Composite Image Generation from Attention Person Regions
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing methods are inefficient and time-consuming for detecting the best shot scene from moving images, as they often include scenes with low importance, poor composition, and low image quality, making it difficult to extract relevant still images.
Innovation Solution
A region detection device and method that extracts still images from moving images, detects the attention person's movement trajectory, and generates a composite image including the detected entire region of the attention person, using units for still image extraction, attention person detection, movement trajectory analysis, and composite image generation.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If manual detection of best shot scenes is performed, then detection accuracy can be maintained, but time consumption and effort increase significantly
Solution Approach 1:
The system enables automatic detection of best shot scenes through self-service mechanisms. The evaluation unit automatically calculates evaluation values based on motion analysis, face detection, and composition assessment, eliminating the need for manual intervention while maintaining detection accuracy through multi-criteria automated evaluation
Solution Approach 2:
The patent replaces manual mechanical detection with an automated computational system. The evaluation unit substitutes human judgment with algorithmic analysis that processes motion trajectories, face regions, and composition metrics to automatically determine best shot scenes, significantly reducing time consumption while preserving detection quality
2Quantity of substance
If all scenes in moving images are extracted as still images, then completeness is improved, but image quality and importance decrease due to inclusion of low-quality scenes
Solution Approach 1:
The patent applies local quality assessment by evaluating different regions and characteristics of each scene independently. The evaluation unit analyzes motion trajectories, face regions, and composition metrics locally for each extracted frame, assigning evaluation values based on local quality indicators rather than uniform extraction, thereby filtering out low-quality scenes while maintaining overall completeness
3Productivity
If simple extraction methods are used, then processing speed is improved, but detection precision deteriorates due to inability to identify important scenes
Solution Approach 1:
The patent segments the detection process into multiple independent evaluation components: motion analysis unit, face region detection unit, composition evaluation unit, and overall evaluation unit. Each segment processes specific aspects independently and efficiently, then combines results to achieve comprehensive detection precision without sacrificing processing speed through parallel computation
Solution Approach 2:
The system performs preliminary actions by pre-calculating and storing motion trajectories, face regions, and composition metrics for each frame before final evaluation. This preliminary processing enables rapid retrieval and combination of features during the evaluation phase, maintaining high processing speed while ensuring detection precision through comprehensive pre-analyzed data
Data Source
AI summary
The image processing apparatus includes a region detection unit that detects a face region of the attention person, an attention person movement region of the moving image, the entire region of the attention person, and an attention person transfer region of the moving image, a region image extraction unit that extracts an image of the face region of the attention person, an image of the attention person movement region of the moving image, an image of the entire region of the attention person, and an image of the attention person transfer region of the moving image, which respectively correspond to the face region of the attention person, the attention person movement region of the moving image, the entire region of the attention person, and the attention person transfer region of the moving image, from the still image, and a composite image generation unit that generates a composite image.


