A three-dimensional multi-modal interaction input determination method and related products

By constructing a multimodal interaction method that combines 3D motion sensing and voice input, and combining it with the context information of the interactive interface, a clear set of interactive command fields is generated. This solves the problem that the execution results in existing 3D interactive tools do not match the user's expectations, and improves the user experience.

CN122431535APending Publication Date: 2026-07-21BEIJING REALSENSE BYTE TECHNOLOGY CO LTD
0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
BEIJING REALSENSE BYTE TECHNOLOGY CO LTD
Filing Date
2026-05-27
Publication Date
2026-07-21

AI Technical Summary

Technical Problem

Existing 3D interactive tools suffer from fragmented and uncommon interaction modes in virtual character creation, partial character editing, and game scene construction. This leads to results that do not meet user expectations, resulting in semantic confusion of input and deviations in command recognition, thus reducing the user experience.

Method used

By constructing a set of spatial interaction parameters based on three-dimensional haptic input and a set of semantic intents based on voice input, and combining the contextual representation information of the current interactive interface, the interpretation constraints are determined, semantic rule mapping and legality filtering are performed, an initial set of control and spatial operation instruction fields is generated, and confidence calculation and structuring are performed to obtain the target set of interaction instruction fields.

Benefits of technology

It effectively reduces the deviation between the execution result and the user's expectations, improves the user experience, and fully restores the user's operation requirements through multimodal information collection, avoiding misidentification and invalid commands, and ensuring that the system's recognized intent is consistent with the user's intent.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN122431535A_ABST
    Figure CN122431535A_ABST
Patent Text Reader

Abstract

The application discloses a three-dimensional multi-modal interaction input determination method and related products. In the scheme, a spatial interaction parameter set is constructed based on the three-dimensional somatosensory input of a user; a semantic intention set is constructed based on the voice input of the user; the state of a current interaction interface is identified to obtain a context representation information set; an interpretation constraint condition is determined based on the context representation information set, and the semantic intention set and the spatial interaction parameter set are processed respectively under the interpretation constraint condition; and a target interaction instruction field set is determined based on candidate object confidence calculation results and candidate interaction action confidence calculation results. The three-dimensional somatosensory input, the voice input and the current interface state input are hierarchically and cooperatively analyzed, and the target interaction instruction field set is output, so that the problems of input semantic confusion and instruction recognition deviation in the prior art are solved, the deviation between the execution result and the user's expectation is reduced, and the user's use experience is improved.
Need to check novelty before this filing date? Find Prior Art