Avatar Generation Using Pre-Processed Visual Artifacts

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing avatar generation systems inaccurately represent users, require high-performance hardware, and struggle on power-constrained devices like smartphones and tablets.

Innovation Solution

Capture sensor data during enrollment, supplemented with user's digital assets to generate accurate virtual representations, using persona and visual artifact networks to enhance geometric and texture characteristics.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If existing avatar generation systems are used, then avatar generation can be performed, but the avatar inaccurately represents the user and requires high-performance hardware

Engineering Contradiction:
Improveavatar accuracyVSAvoidhardware requirements
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The system performs preliminary actions by collecting multiple media assets (photos, videos) of the user in advance during an enrollment phase. These assets are processed offline to create a comprehensive library of user-specific visual artifacts, which are then stored for later use. This preliminary collection and processing enables accurate avatar generation without requiring high computational power during the actual avatar creation moment.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent creates an equipotential solution by providing consistent, high-quality avatar generation capability across devices with different computational powers. By pre-processing data on any device during enrollment and storing the results, the system ensures that both high-performance and low-performance devices can generate accurate avatars using the same pre-computed user-specific artifacts, equalizing the capability across different hardware levels.

Inventive Principle:
Principle #12Equipotentiality

2Ease of manufacture

If sensor data captured at a particular time is used, then avatar generation is simple, but the avatar does not accurately represent the user's typical appearance

Engineering Contradiction:
Improveavatar generation simplicityVSAvoiduser representation accuracy
Core Design Contradiction:
Ease of manufactureVSMeasurement precision

Solution Approach 1:

The system merges multiple data sources - combining sensor data captured during enrollment with pre-collected media assets (photos, videos) from the user's device. By integrating these diverse sources of user appearance information, the system creates a comprehensive representation that captures the user's typical appearance across different conditions, lighting, and time periods, rather than relying on a single sensor capture.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The patent applies parameter changes by transforming the static, single-time sensor data into a dynamic, multi-temporal representation. The system varies the temporal parameter by incorporating media assets from different times, and transforms the data format by extracting visual artifacts (geometric and texture parameters) from multiple sources, creating a more robust and accurate user representation that accounts for appearance variations over time.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS20250111611A1Generating Virtual Representations Using Media Assets
Publication Date: 2025.04.03 APPLE INC
  • US20250111611A1 patent drawing
  • US20250111611A1 patent drawing
  • US20250111611A1 patent drawing

AI summary

Generating a 3D representation of a subject includes obtaining sensor data of a subject. Media assets comprising the subject can be obtained from a digital asset library. A visual artifact for the subject can be generated from the media assets and used, along with the sensor data, to generate one or more virtual representations of the subject. Visual artifacts include textural and/or geometric characteristics of the visual appearance of the subject and are derived from image data in the media assets.