Live Image Capture With Palm Gesture Object Tracking
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Users need to control image capturing devices during live streaming without assistance to track objects accurately and switch display modes quickly, such as determining object location, size, and switching between tracking and normal modes.
Innovation Solution
An image capturing device with an image processing unit that analyzes video images using neural networks for object tracking, palm gesture recognition, and adaptive display modes like picture-in-picture and side-by-side, allowing automatic object tracking and mode switching based on user gestures.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If the user manually controls the image capturing device during live streaming, then the user can introduce products or give speeches, but the user cannot simultaneously control the device and perform other actions without assistance
Solution Approach 1:
The system performs automatic object tracking and display mode switching without requiring user intervention. The control unit automatically determines object locations, calculates tracking parameters, and switches between display modes based on gesture recognition, enabling the device to serve itself and eliminate the need for assistant operators
Solution Approach 2:
The patent replaces manual mechanical control with automated electronic control systems. The control unit uses gesture recognition and automatic algorithms to substitute human manual operations, transforming the control mechanism from mechanical/manual to electronic/automated
2Measurement precision
If the image capturing device tracks objects automatically, then the tracking accuracy is improved, but the processing time and computational load increase
Solution Approach 1:
The system performs preliminary actions by pre-calculating tracking parameters and preparing display mode transitions in advance. The control unit determines object locations and calculates tracking parameters proactively, so that when tracking is needed, the information is already prepared, reducing actual processing time
Solution Approach 2:
The system maintains continuous object tracking and continuous monitoring of gesture inputs, eliminating the need to stop and restart processing. The control unit continuously calculates tracking parameters and continuously recognizes gestures, ensuring uninterrupted useful action and reducing overall processing time
3Adaptability or versatility
If the device switches between tracking mode and normal mode, then the adaptability to different scenarios is improved, but the complexity of mode management increases
Solution Approach 1:
The system dynamically switches between tracking mode and normal mode based on real-time gesture recognition. The control unit automatically transitions between modes without requiring manual configuration, making the system adaptive to different scenarios while keeping mode management simple through automatic detection and switching
4Speed
If the user needs to quickly control the image capturing device to determine object location and size, then the response speed is improved, but the operational complexity increases
Solution Approach 1:
The control unit automatically determines object locations and calculates size parameters without requiring user input. The system performs these operations autonomously by processing video images and generating tracking parameters, achieving fast response speed while maintaining operational simplicity
Data Source
AI summary
An image capturing method comprising: obtaining a plurality of video images; analyzing whether there is a palm in the video images and identifying a palm gesture; when the palm gesture is a tracking gesture, entering a tracking identification mode, so as to use an interaction manner between a user and an object to determine that the object is a tracking object, and calculating relevant information of the tracking object; tracking the tracking object using a tracking operation, so as to generate a plurality of tracking images; and according to a first video display mode, using the tracking images and the video images to generate a plurality of live images.


