Adaptive Feature Encoding And Decoding For AI Image Compression

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing image compression technologies are unsuitable for artificial intelligence services due to their focus on high-resolution, high-quality image processing for human vision, lacking efficiency and adaptability for machine-oriented tasks.

Innovation Solution

A feature encoding/decoding method and apparatus that determines the type and property of the input source, incorporating supplementary and meaning information, and provides privacy protection, enabling efficient encoding/decoding of feature data for machine tasks.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If existing image compression technology is used, then high-resolution and high-quality image processing for human vision is achieved, but it is unsuitable for artificial intelligence services

Engineering Contradiction:
Improveimage qualityVSAvoidsuitability for AI services
Core Design Contradiction:
Measurement precisionVSAdaptability or versatility

Solution Approach 1:

The patent introduces dynamic adaptation by detecting whether the input is natural image data or feature data and switching encoding modes accordingly. The encoding apparatus dynamically adjusts its operation based on input type identification, enabling it to adapt between human-vision-oriented and machine-task-oriented processing modes.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent changes the fundamental parameters of the encoding process by introducing feature map type information (first feature map type, second feature map type) and adjusting encoding strategies based on these parameters. This allows the system to optimize for different AI service requirements while maintaining compatibility with traditional image processing.

Inventive Principle:
Principle #35Parameter changes

2Measurement precision

If image compression technology optimized for human vision is used, then high-quality images are produced, but encoding/decoding efficiency for machine tasks is reduced

Engineering Contradiction:
Improveimage qualityVSAvoidencoding/decoding efficiency
Core Design Contradiction:
Measurement precisionVSProductivity

Solution Approach 1:

The patent segments the encoding process into distinct pathways based on input type. By dividing the encoding logic into natural image encoding and feature data encoding branches, the system can apply optimized strategies for each type, improving overall efficiency for machine tasks while preserving quality for human vision applications.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces an intermediary mechanism (input type detection and feature map type identification) that mediates between the encoding apparatus and the input data. This intermediary layer enables the system to select appropriate encoding strategies, thereby improving productivity for AI services without sacrificing image quality when needed.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Productivity

If feature data encoding is implemented, then encoding efficiency for AI services is improved, but complexity of determining input type and feature map type increases

Engineering Contradiction:
Improveencoding efficiencyVSAvoidinput type determination complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent applies preliminary action by performing input type detection and feature map type identification before the actual encoding process. By determining the input type in advance and preparing the appropriate encoding strategy beforehand, the system reduces complexity during the main encoding operation and improves overall efficiency.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The encoding apparatus performs self-service by automatically detecting input types and selecting appropriate encoding modes without requiring external intervention. The system self-determines whether to process natural images or feature data and adjusts its operation accordingly, managing the complexity internally while maintaining simplicity for users.

Inventive Principle:
Principle #25Self-service

4Measurement precision

If feature map type information is encoded, then decoding accuracy is improved, but amount of encoding information increases

Engineering Contradiction:
Improvedecoding accuracyVSAvoidencoding information volume
Core Design Contradiction:
Measurement precisionVSQuantity of substance

Solution Approach 1:

The patent applies partial action by selectively encoding feature map type information only when necessary for accurate decoding. Rather than encoding all possible information, the system encodes only the essential feature map type identifiers (first or second feature map type) that are critical for maintaining decoding accuracy, thereby balancing information volume with precision.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS20250234005A1Feature encoding/decoding method, device, recording medium storing bitstream, and method for transmitting bitstream
Publication Date: 2025.07.17 LG ELECTRONICS INC
  • US20250234005A1 patent drawing
  • US20250234005A1 patent drawing
  • US20250234005A1 patent drawing

AI summary

Provided are a feature encoding/decoding method and device, a recording medium on which a bitstream generated by the feature encoding method is stored, and a method for transmitting the bitstream. The feature decoding method according to the present disclosure comprises the steps of: determining whether or not an input source is feature data; determining a feature map type of the feature data based on the input source being the feature data; obtaining encoding information of the feature data from a bitstream; and decoding the feature data based on the feature map type and the encoding information, and whether or not the input source is the feature data may be determined based on input format identification information that is obtained from a sequence parameter set.