Large Model Video Processing for 3D Posture Training Assessment

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing physical training methods lack standardized assessment and guidance for users, leading to inconsistent training quality and effectiveness.

Innovation Solution

A large model-based video processing method that collects imitation videos, extracts three-dimensional postures, and performs comparative assessments using pre-trained models to provide accurate feedback and suggestions for improvement.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If users train independently without standardized assessment, then training flexibility is improved, but training quality consistency deteriorates

Engineering Contradiction:
Improvetraining flexibilityVSAvoidtraining quality consistency
Core Design Contradiction:
Ease of operationVSManufacturing precision

Solution Approach 1:

The system provides automated feedback by comparing user imitation videos against standard videos through three-dimensional posture extraction and assessment. The large model evaluates posture accuracy, movement standardization, and training effectiveness, delivering structured feedback reports that guide users to improve their training quality while maintaining independent training flexibility.

Inventive Principle:
Principle #23Feedback

Solution Approach 2:

The patent replaces manual mechanical assessment by personal trainers with an automated computer vision system. The large model processes videos, extracts three-dimensional postures, and performs objective evaluation without human intervention, ensuring consistent and standardized training quality assessment while users maintain training flexibility.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Manufacturing precision

If personal trainers provide guidance and assessment, then training quality is improved, but accessibility and cost deteriorate

Engineering Contradiction:
Improvetraining qualityVSAvoidaccessibility
Core Design Contradiction:
Manufacturing precisionVSEase of operation

Solution Approach 1:

The system enables self-service training assessment where users independently upload their workout videos and receive automated evaluation from the large model. The system extracts three-dimensional postures, compares them against standard videos, and provides detailed feedback reports without requiring users to hire or visit personal trainers, significantly improving accessibility while maintaining training quality.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent substitutes the mechanical service of personal trainers with an automated AI-based assessment system. The large model performs posture extraction, comparison, and evaluation functions that previously required human experts, making high-quality training assessment accessible to all users regardless of geographic location or financial constraints.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

3Speed

If traditional video processing methods are used, then processing speed is improved, but assessment accuracy deteriorates

Engineering Contradiction:
Improveprocessing speedVSAvoidassessment accuracy
Core Design Contradiction:
SpeedVSMeasurement precision

Solution Approach 1:

The system changes the fundamental parameters of video processing by transitioning from traditional two-dimensional image analysis to three-dimensional posture extraction using large models. This parameter change enables the system to capture spatial relationships, joint angles, and movement trajectories with high precision, significantly improving assessment accuracy while maintaining efficient processing speeds through optimized model architecture.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS20250218037A1Large model-based video processing method, device and storage medium
Publication Date: 2025.07.03 BEIJING BAIDU NETCOM SCI & TECH CO LTD
  • US20250218037A1 patent drawing
  • US20250218037A1 patent drawing
  • US20250218037A1 patent drawing

AI summary

A large model-based video processing method, device and storage medium in the field of artificial intelligence technology, particularly in the fields of deep learning and large models are disclosed. The specific solution includes: collecting an imitation video made by a user based on a target video; extracting three-dimensional postures of the imitation video using a pre-trained large model based on the imitation video; and performing posture assessment on the imitation video using the pre-trained large model based on the three-dimensional postures of the imitation video and pre-obtained three-dimensional postures of the target video to obtain an assessment result.