Video Editing Template with Face Detection and Shot Classification

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Amateur movie makers face challenges in editing voluminous video footage into engaging and professional-looking presentations due to the requirement of significant skill, experience, and creativity, often resulting in unedited or poorly edited videos that are long, disjointed, and boring.

Innovation Solution

A method that analyzes video clips to identify frames with faces, classifies them based on face-related characteristics, and presents them in a user interface for selection, allowing users to generate movies using a template with predefined shot placeholders, background music, and effects, thereby simplifying the editing process.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of manufacture

If amateur movie makers manually edit voluminous video footage, then they can create customized presentations, but the process requires significant skill, experience, effort and creativity, resulting in long and disjointed videos

Engineering Contradiction:
Improveease of video editingVSAvoidediting efficiency
Core Design Contradiction:
Ease of manufactureVSProductivity

Solution Approach 1:

The system automatically analyzes video clips to detect faces and classify shots, enabling the editing process to serve itself without requiring expert intervention. The automated face detection and shot classification mechanisms allow the software to independently organize and structure video content, freeing users from manual editing tasks while maintaining professional-quality output

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The system performs preliminary analysis of all video clips before the user begins editing, pre-detecting faces and pre-classifying shots into categories such as close-up, medium, and wide shots. This preliminary organization of video content based on face detection results allows users to quickly assemble presentations without having to manually analyze each clip, significantly improving editing efficiency

Inventive Principle:
Principle #10Preliminary action

2Manufacturing precision

If professional editing techniques are used, then video quality improves, but the complexity of the editing process increases significantly

Engineering Contradiction:
Improvevideo editing qualityVSAvoidediting process complexity
Core Design Contradiction:
Manufacturing precisionVSDevice complexity

Solution Approach 1:

The system introduces an intermediary layer of automated shot classification that translates raw video footage into structured, categorized content. By detecting faces and classifying shots into standard cinematic categories (close-up, medium, wide), the system bridges the gap between amateur footage and professional editing requirements, enabling high-quality output through a simplified interface that masks the underlying complexity

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The system changes the parameter of shot classification from manual user judgment to automated face detection-based classification. By analyzing facial features and their position in frames, the system automatically assigns shots to appropriate categories, maintaining professional editing standards while eliminating the need for users to understand complex editing parameters and techniques

Inventive Principle:
Principle #35Parameter changes

3Ease of operation

If automated face detection is implemented, then video clip organization improves, but analysis time and processing requirements increase

Engineering Contradiction:
Improvevideo clip organizationVSAvoidanalysis time
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The system applies partial action by performing face detection on representative frames rather than every single frame in each video clip. This selective sampling approach provides sufficient information for accurate shot classification while significantly reducing processing time and computational resources required, balancing organization quality with time efficiency

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS8726161B2Visual presentation composition
Publication Date: 2014.05.13 APPLE INC
  • US8726161B2 patent drawing
  • US8726161B2 patent drawing
  • US8726161B2 patent drawing

AI summary

Methods, systems and/or computer program products are disclosed that help facilitate visual presentation composition. A method includes analyzing a plurality of video clips, each video clip comprising a plurality of frames, to determine a subset of the plurality of video clips that have at least one frame depicting one or more faces. The method further includes presenting, in a user interface of a video editing application, the determined subset of video clips along with indicia indicating one or more face-related characteristics of each of the subset of video clips. Furthermore, the method includes receiving, from a user of the video editing application, a selection of one or more frames of at least one of the subset of video clips to populate a shot placeholder in a movie-building template, and generating a playable media file representing a movie based at least in part on the selection received from the user.