Encrypted MPEG-2 Splicing via Pre-conditioned Cue Messages

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Traditional MPEG splicing techniques require PES packets to be in the clear, making it impossible to splice encrypted MPEG streams, leading to issues like lip sync problems and decoder buffer model violations.

Innovation Solution

A pre-conditioning encoder inserts SCTE-35 cue messages and encodes video and audio streams to create Random Access Points, allowing seamless splicing of encrypted MPEG-2 Transport Streams by aligning audio and video frames within PES packets, maintaining decoder buffer compliance.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If traditional splicing techniques are used on encrypted MPEG streams, then security is improved, but splicing capability deteriorates (cannot splice)

Engineering Contradiction:
ImprovesecurityVSAvoidsplicing capability
Core Design Contradiction:
ReliabilityVSAdaptability or versatility

Solution Approach 1:

The encoder performs pre-conditioning actions before encryption by inserting SCTE-35 cue messages and creating Random Access Points (RAPs) at specific locations in the bitstream. This preliminary structuring allows the splicer to identify valid splice points and perform seamless splicing on encrypted streams without needing to decrypt the content first.

Inventive Principle:
Principle #10Preliminary action

2Reliability

If splicing is performed on encrypted streams without pre-conditioning, then security is maintained, but audio-video synchronization deteriorates (lip sync problems)

Engineering Contradiction:
ImprovesecurityVSAvoidaudio-video synchronization
Core Design Contradiction:
ReliabilityVSManufacturing precision

Solution Approach 1:

The encoder pre-conditions the audio and video streams by aligning their frame structures and inserting synchronization markers (SCTE-35 cue messages) before encryption. This ensures that when splicing occurs on the encrypted stream, the audio and video frames remain properly synchronized without causing lip sync issues.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The encoder modifies the temporal parameters of audio and video frames by adjusting their presentation timestamps (PTS) and decoding timestamps (DTS) to ensure proper alignment. This parameter adjustment maintains audio-video synchronization throughout the splicing process even when content is encrypted.

Inventive Principle:
Principle #35Parameter changes

3Reliability

If traditional splicing modifies frame sizes to comply with decoder buffer model, then decoder compliance is improved, but stream integrity deteriorates (requires clear content)

Engineering Contradiction:
Improvedecoder complianceVSAvoidstream integrity
Core Design Contradiction:
ReliabilityVSLoss of information

Solution Approach 1:

The encoder pre-conditions the stream by inserting RAPs and adjusting decoder buffer parameters before encryption. This allows the splicer to maintain decoder buffer compliance through timestamp adjustments rather than frame modification, preserving the integrity of the encrypted content.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The invention replaces the traditional mechanical approach of modifying frame sizes and content with a signal-based approach using SCTE-35 cue messages and timestamp adjustments. This substitution allows compliance with the decoder buffer model without altering the actual encrypted video frames, maintaining stream integrity.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Data Source

PatentUS8781003B2Splicing of encrypted video/audio content
Publication Date: 2014.07.15 TRITON US VP ACQUISITION CO
  • US8781003B2 patent drawing
  • US8781003B2 patent drawing
  • US8781003B2 patent drawing

AI summary

System and method for performing a splice operation on an encrypted or unencrypted MPEG-2 transport stream. A splice trigger is received at a pre-conditioning encoder. In response, the encoder generates, e.g., an SCTE-35 cue message that is intended to be received by a splicer. Also in response to the splice trigger, the encoder encodes/conditions a network feed such that a decoder buffer delay reaches a predefined value at a video frame of the network feed that corresponds to a splice point indicated by the SCTE-35 cue message. The network feed may then be encrypted in a known fashion. At the splicer, another feed is switched into the stream at the splice point, wherein the another feed is encoded such that a decoder buffer delay at a video frame of the another feed corresponding to the splice point is the same as the predefined value. The predefined value is defined as DTS-STC, where DTS is a Decoding Time Stamp and STC is a System Time Clock. The pre-conditioning encoder generates around the splice point audio PES packets that contain a single aligned audio frame.