Video Call Gesture Recognition Through Modular Image Processing
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing electronic devices lack efficient methods for recognizing user gestures during video calls, limiting their functionality and user interaction capabilities.
Innovation Solution
An electronic device equipped with a camera module, communication module, and display module, capable of identifying a video call reception event, capturing images, and performing operations based on recognized gestures through a gesture recognition process involving a processor and AI model.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If gesture recognition is added to video call functionality, then user interaction capability is improved, but device complexity increases
Solution Approach 1:
The gesture recognition system is segmented into distinct modules: camera module for image capture, processor for image processing and gesture identification, and communication module for video call handling. This modular segmentation allows each component to perform its specific function independently, improving user interaction capability while managing device complexity through functional division.
2Speed
If real-time gesture recognition is implemented during video calls, then responsiveness is improved, but processing time and energy consumption increase
Solution Approach 1:
The system performs preliminary actions by continuously capturing images through the camera module during video calls and pre-processing these images to identify gestures. By preparing gesture data in advance and maintaining a buffer of processed image information, the system achieves real-time responsiveness without intensive processing during critical moments, thereby reducing peak energy consumption while maintaining speed.
Data Source
AI summary
An electronic device is provided. The electronic device includes memory storing instructions, a camera module, a communication module, a display module, and at least one processor operatively connected to the memory, the camera module, the communication module, and the display module. The instructions, that when executed by the at least one processor, cause the electronic device to identify a reception event of a video call based on the communication module, hook one or more images captured based on the camera module to provide image information corresponding to the hooked one or more images to a first application, identify a gesture based on information output from the first application, by inputting the image information to the first application, and perform at least one operation corresponding to the identified gesture.


