Agentless KVM OCR Text Capture From Remote Video Frames

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing agentless KVM systems cannot extract and process text or alphanumeric information from video frames for further use, necessitating the use of agent-based solutions that are undesirable due to security and complexity concerns.

Innovation Solution

Implementing an optical character recognition (OCR) software application on the client computing device to convert selected text or alphanumeric information from a video frame into ASCII text output and copy it to the clipboard for subsequent use.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If agentless KVM system is used to avoid software installation on target system, then security and complexity are improved, but text extraction capability deteriorates

Engineering Contradiction:
ImprovesecurityVSAvoidtext extraction capability
Core Design Contradiction:
ReliabilityVSAdaptability or versatility

Solution Approach 1:

The patent introduces an intermediary OCR processing layer between the agentless KVM system and the user. The KVM appliance captures video frames from the target system and transmits them to the client device, where OCR software processes the video frames to extract text. This intermediary approach enables text extraction without requiring agents on the target system, thus maintaining security while adding functionality.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent replaces the traditional mechanical approach of installing text extraction agents on target systems with an optical-based solution. Video frames captured by the agentless KVM system are processed through OCR (optical character recognition) software on the client device, substituting the need for mechanical software installation on the target system with an optical processing approach on the client side.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Device complexity

If agentless KVM system is used, then device complexity is reduced, but text processing functionality is lost

Engineering Contradiction:
Improvesystem complexityVSAvoidtext processing functionality
Core Design Contradiction:
Device complexityVSAdaptability or versatility

Solution Approach 1:

The patent introduces an intermediary OCR processing layer between the agentless KVM system and the user. The KVM appliance captures video frames from the target system and transmits them to the client device, where OCR software processes the video frames to extract text. This intermediary approach enables text extraction without requiring agents on the target system, thus maintaining security while adding functionality.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent shifts the text processing functionality from the target system dimension to the client device dimension. Instead of processing text directly on the target system or requiring agents there, the solution moves the processing capability to another dimension - the client device where video frames are received and processed through OCR, thereby maintaining system simplicity while adding functionality.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

3Ease of operation

If traditional hardware-based KVM is used, then agentless operation is achieved, but text selection and copying capability deteriorates

Engineering Contradiction:
Improveagentless operationVSAvoidtext selection and copying capability
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The patent replaces the traditional mechanical approach of installing text extraction agents on target systems with an optical-based solution. Video frames captured by the agentless KVM system are processed through OCR (optical character recognition) software on the client device, substituting the need for mechanical software installation on the target system with an optical processing approach on the client side.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The patent introduces an intermediary OCR processing layer between the agentless KVM system and the user. The KVM appliance captures video frames from the target system and transmits them to the client device, where OCR software processes the video frames to extract text. This intermediary approach enables text extraction without requiring agents on the target system, thus maintaining security while adding functionality.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentEP4296838B1System and method for OCR-based text conversion and copying mechanism for agentless hardware-based KVM
Publication Date: 2026.04.08 VERTIV CORP
  • EP4296838B1 patent drawingFigure 1
  • EP4296838B1 patent drawingFigure 2

AI summary

The present disclosure relates to a method for selecting and copying one or more characters of at least one of text or alphanumeric information appearing within a video image frame being displayed on a display of a client computing device, during a keyboard, video and mouse (KVM) session with a remote KVM appliance. The method enables a user to define text or alphanumeric information being displayed in a video frame on the display, using a control component of the client computing device, which the user desires to convert into text. The method uses an optical character recognition (OCR) software application to convert the selected video information into a text output. The text output can then be copied and pasted into one or more other applications, documents or web pages by the user for subsequent use.