Output Layer Sets for Spatial/SNR Video Layer Selection

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing video coding systems require decoders to always support the highest encoded layer, leading to scalability issues as they cannot adapt to intermediate layers based on hardware and network requirements, resulting in errors and inefficient resource utilization.

Innovation Solution

The use of Output Layer Sets (OLSs) to support spatial and SNR scalability, where an ols_mode_idc syntax element in the Video Parameter Set (VPS) indicates the number of OLSs, with each OLS containing specific layers, allowing the decoder to determine the output layer quickly and decode only the necessary layers.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Manufacturing precision

If decoders are configured to always support the highest encoded layer, then video quality is maximized, but system adaptability to intermediate layers deteriorates

Engineering Contradiction:
Improvevideo qualityVSAvoidadaptability to intermediate layers
Core Design Contradiction:
Manufacturing precisionVSAdaptability or versatility

Solution Approach 1:

The patent segments the video data into multiple independent layers with different quality levels. Each layer can be decoded independently or in combination with other layers, allowing decoders to select appropriate layers based on hardware capabilities and network conditions. This segmentation enables intermediate layer decoding without requiring support for the highest layer.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces dynamic layer selection mechanisms where decoders can adaptively choose which layers to decode based on real-time hardware capabilities and network conditions. The system dynamically adjusts the decoding configuration to match available resources, enabling flexible adaptation to intermediate layers while maintaining optimal video quality when resources permit.

Inventive Principle:
Principle #15Dynamics

2Adaptability or versatility

If multiple layers are encoded to support different quality levels, then scalability is improved, but bitstream complexity increases

Engineering Contradiction:
ImprovescalabilityVSAvoidbitstream complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent merges multiple layers into a unified bitstream structure with standardized syntax elements. The VPS (Video Parameter Set) contains consolidated layer configuration information, and OLS (Output Layer Set) structures group layers systematically. This merging reduces bitstream complexity by eliminating redundant signaling while maintaining scalability across different quality levels.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The patent uses parameter-based layer differentiation where layers are defined by specific parameter sets (resolution, SNR, frame rate) rather than complex structural variations. The ols_mode_idc parameter and related syntax elements provide compact signaling that changes parameters efficiently to indicate layer configurations, reducing bitstream overhead while supporting multiple quality levels.

Inventive Principle:
Principle #35Parameter changes

3Manufacturing precision

If all encoded layers are decoded, then maximum video quality is achieved, but resource utilization efficiency deteriorates

Engineering Contradiction:
Improvevideo qualityVSAvoidresource utilization efficiency
Core Design Contradiction:
Manufacturing precisionVSLoss of energy

Solution Approach 1:

The patent extracts and processes only the necessary layers required for the desired output quality. Instead of decoding all encoded layers, the system selectively extracts the appropriate layer or combination of layers needed to meet quality requirements. This extraction approach reduces computational resources and energy consumption while maintaining adequate video quality.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent implements partial decoding where only the necessary portion of the multi-layer bitstream is processed. Decoders can perform partial action by decoding only up to the required quality level without processing higher quality layers. This partial action approach optimizes resource utilization by avoiding unnecessary decoding operations while still achieving the target video quality.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS20250294162A1OLS For Spatial And SNR Scalability
Publication Date: 2025.09.18 HUAWEI TECH CO LTD
  • US20250294162A1 patent drawing
  • US20250294162A1 patent drawing
  • US20250294162A1 patent drawing

AI summary

A video coding mechanism is disclosed. The mechanism includes encoding a bitstream comprising one or more layers of coded pictures. A video parameter set (VPS) is also encoded into the bitstream. The VPS includes an output layer set (OLS) mode identification code (ols_mode_idc) specifying that a total number of OLSs specified by the VPS is equal to a number of layers specified by the VPS. The bitstream is stored for communication toward a decoder.