Head Avatar Generation From Shot Video With Neural Implicit Modeling

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional avatar creation technology requires expertise or complex systems, leading to reduced graphic precision and difficulty in modifying avatars.

Innovation Solution

A method and device for creating a head avatar using a smartphone camera, involving neural implicit functions (NIF) to generate precise head meshes and textures from shot videos, optimizing differences between input images and rendering images, and allowing user-modifiable avatars.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Manufacturing precision

If conventional avatar creation technology is used, then expertise or complex systems are required, but this leads to reduced graphic precision and difficulty in modifying created avatars

Engineering Contradiction:
Improvegraphic precisionVSAvoidsystem complexity
Core Design Contradiction:
Manufacturing precisionVSDevice complexity

Solution Approach 1:

The patent replaces conventional mechanical/complex system-based avatar creation with a neural network-based automated system. The neural network learns from input images to generate 3D head models, substituting the need for expert operators and complex manual systems with an intelligent automated system that achieves higher precision while reducing operational complexity

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The patent uses 2D input images as copies or representations of the target subject's face, and through neural network processing, generates accurate 3D head model copies. This copying approach allows the system to capture precise facial features from simple images without requiring complex scanning equipment or expert intervention

Inventive Principle:
Principle #26Copying

2Adaptability or versatility

If conventional avatar creation technology is used, then complex systems are required, but this leads to difficulty in modifying created avatars

Engineering Contradiction:
Improveavatar modification capabilityVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent segments the avatar creation process into distinct components: input image processing, neural network learning, 3D mesh generation, and texture mapping. This segmentation allows each component to be independently optimized and modified, enhancing adaptability while maintaining system manageability and reducing overall complexity

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent creates a dynamic avatar system where the neural network can be retrained with new input images, and the generated 3D models can be modified by adjusting network parameters or input data. This dynamic nature allows easy adaptation and modification without requiring complex reconfiguration of the entire system

Inventive Principle:
Principle #15Dynamics

Data Source

PatentUS12614364B2Method and device for creating head avatar using shot video
Publication Date: 2026.04.28 KOREA INST OF SCI & TECH
  • US12614364B2 patent drawing
  • US12614364B2 patent drawing
  • US12614364B2 patent drawing

AI summary

According to a method and device for generating a head avatar using a shot video, a head mesh of neutral look and a head texture map of neutral look are generated based on a base mesh of a base model and at least one shooting image representing a face of neutral look among a plurality of shooting images generated from a camera shot video, and a head mesh reflecting a certain facial expression and a head texture reflecting the certain facial expression are generated based on at least one shooting image representing a face of a certain facial expression among the plurality of shooting images, the head mesh of neutral look, and the head texture map of neutral look.