Metaverse Avatar Head Overlay for Real-Time Video Matching
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current metaverse-based virtual office environments struggle to convey nuanced information like facial expressions and actions beyond text, limiting social interactions and bond formation among users, as they rely on static images rather than real-time video feeds.
Innovation Solution
A method that identifies and matches real-time user image data from a video camera with a 3D avatar in a metaverse-based office environment, projecting the image as a two-dimensional shape at the avatar's head position, allowing for the conveyance of facial expressions and actions within a chat group while preventing information leakage by restricting conversation to specific groups.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If real-time video camera image data is received and overlapped with 3D avatar in metaverse-based office environment, then social ties and communication accuracy are improved by conveying facial expressions and actions, but device complexity and processing requirements increase
Solution Approach 1:
The patent creates a virtual copy of the user's real-world image by capturing video camera data and overlaying it onto the 3D avatar in the metaverse environment. This copying approach preserves the original facial expressions and actions while transferring them to the virtual space, enabling accurate communication without requiring complex real-time rendering of physical features
Solution Approach 2:
The patent transitions from two-dimensional video images to three-dimensional virtual space by overlaying the captured image onto the 3D avatar. This dimensional transformation allows the flat video feed to be integrated into the immersive 3D metaverse environment, maintaining facial expression accuracy while adapting to the virtual dimension
2Reliability
If user image data is placed at avatar head position in 3D virtual space, then communication realism is improved by replacing character face with real-time user image, but processing time and computational resources increase
Solution Approach 1:
The patent performs preliminary actions by capturing and preparing the user's image data before overlaying it onto the avatar. The system pre-processes the video feed and positions it at the avatar's head location in advance, reducing real-time processing requirements during actual communication interactions
Solution Approach 2:
The patent introduces an intermediary layer by overlaying the real-time video image onto the 3D avatar rather than directly replacing or rendering complex facial features. This intermediary approach simplifies processing by using the captured image as a mediator between the user's physical appearance and the virtual avatar representation
3Object-affected harmful factors
If chat group identification and user image matching are implemented, then information security is improved by preventing information leakage, but device complexity and operational complexity increase
Solution Approach 1:
The patent segments the metaverse environment into distinct chat groups with separate communication channels. By dividing users into specific groups and restricting image data overlay to only those within the same chat group, the system prevents information leakage while maintaining ease of operation through automated group identification and access control
Data Source
AI summary
The present invention relates to a method for user image data matching in a metaverse-based office environment and, more particularly, to a method for user image data matching in a metaverse-based office environment, the method comprising: a chat group identification step of identifying a camera viewpoint of a target user in a virtual space, and identifying whether a chat group for users included in a virtual image of the camera viewpoint is included; and a user image matching step of matching user images to avatars of respective group users included in the chat group, the user images being obtained by image-capturing the group users in real time.


