Video Encoding Apparatus Frame-Level Data Extraction

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current video encoding methods, such as the H.264 standard, are not suitable for real-time stream-typed video traffic as they require large granularity encoded picture data, which is not applicable for immediate frame-level data transmission in scenarios like video chatting.

Innovation Solution

A video encoding method and apparatus that processes original picture data to generate and encapsulate multimedia audio and video files in sequence, parsing and modifying data segments to achieve small granularity encoded picture data conforming to standards like H.264, utilizing multiple threads for efficient processing and outputting encoded data with sequence and picture parameter sets.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If video encoding is performed using standard methods (H.264) with large granularity encoded picture data, then encoding compatibility and standard compliance are improved, but real-time stream-typed video traffic transmission capability deteriorates

Engineering Contradiction:
Improveencoding compatibilityVSAvoidreal-time transmission capability
Core Design Contradiction:
Adaptability or versatilityVSProductivity

Solution Approach 1:

The patent segments the encoded picture data by extracting individual frame data from the encoded bitstream. Instead of transmitting complete large-granularity encoded files, the system divides the data into frame-level segments that can be transmitted and processed independently in real-time, resolving the contradiction between standard compliance and real-time transmission capability

Inventive Principle:
Principle #1Segmentation

2Manufacturing precision

If video encoding processes complete video segments before output, then encoding quality and standard compliance are improved, but transmission delay increases

Engineering Contradiction:
Improveencoding qualityVSAvoidtransmission delay
Core Design Contradiction:
Manufacturing precisionVSLoss of time

Solution Approach 1:

The patent performs preliminary encoding processing on individual frames as they are captured, rather than waiting to encode complete video segments. By preparing frame-level encoded data in advance and making it available for immediate transmission, the system reduces transmission delay while maintaining encoding quality through standard-compliant encoding processes

Inventive Principle:
Principle #10Preliminary action

3Productivity

If small granularity frame-level encoded picture data is generated, then real-time stream-typed video traffic transmission is improved, but system compatibility with existing standards deteriorates

Engineering Contradiction:
Improvereal-time transmission capabilityVSAvoidsystem compatibility
Core Design Contradiction:
ProductivityVSAdaptability or versatility

Solution Approach 1:

The patent extracts frame-level encoded picture data from the encoded bitstream by parsing and identifying frame boundaries and data structures. This extraction process retrieves small-granularity frame data while maintaining the standard-compliant encoding format, enabling real-time transmission without sacrificing system compatibility with existing video decoding systems

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS9936266B2Video encoding method and apparatus
Publication Date: 2018.04.03 TENCENT TECHNOLOGY (SHENZHEN) CO LTD
  • US9936266B2 patent drawing
  • US9936266B2 patent drawing
  • US9936266B2 patent drawing

AI summary

A video encoding method and apparatus are provided, in which the method comprises the steps of obtaining respective original picture data in sequence; generating respective multi-media audio and video files in sequence according to the obtained original picture data; parsing each multi-media audio and video file, encapsulating the result of parsing according to a predetermined standard to obtain encoded picture data corresponding to each multi-media audio and video file and conforming to the predetermined standard, and outputting the encoded picture data. The solutions in the present disclosure may meet the need of stream-typed video traffic for small granularity encoded picture data in frame level.