Data Flow Tracking Tool for Distributed Systems

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

In complex distributed computing systems, it is challenging to identify and trace data flows effectively as the systems grow in complexity, making it difficult for administrative users to assess and maintain data validity across applications.

Innovation Solution

A distributed computing environment with a data flow software tool that identifies potential data flows, generates a GUI to visualize upstream and downstream applications, and allows users to verify data flows, facilitating the tracking and remediation of data flows through various functional modules such as data flow discovery, tracing, and visualization.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If the complexity of computing systems increases to handle more applications and data flows, then the system's functionality and processing capability improve, but the difficulty of identifying and tracing data flows increases

Engineering Contradiction:
Improvesystem functionalityVSAvoiddata flow tracing difficulty
Core Design Contradiction:
Adaptability or versatilityVSDifficulty of detecting and measuring

Solution Approach 1:

The patent introduces a data flow tracing system that acts as an intermediary between applications and administrators. This system includes data flow trackers embedded in applications that automatically capture and report data flow information, eliminating the need for administrators to manually trace complex data flows through numerous applications. The intermediary system processes and visualizes data flow paths, making them easily identifiable despite system complexity.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent implements self-service mechanisms where applications automatically generate and report their own data flow information without external intervention. Data flow trackers within applications autonomously monitor, capture, and transmit data flow metadata to the tracing system, enabling automatic data flow identification and tracing without requiring administrators to manually investigate each application's data processing activities.

Inventive Principle:
Principle #25Self-service

2Measurement precision

If administrators manually trace data flows through complex systems, then they can identify data flow paths, but the time and resources required increase significantly

Engineering Contradiction:
Improvedata flow identification accuracyVSAvoidassessment time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The patent implements preliminary action by embedding data flow trackers in applications that continuously and automatically capture data flow information before administrators need to assess it. The system proactively collects, stores, and organizes data flow metadata including source applications, destination applications, and data transformations, so when administrators need to trace data flows, the information is already prepared and readily available, eliminating time-consuming manual tracing.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent establishes feedback mechanisms where the data flow tracing system continuously monitors and reports data flow status to administrators. The system provides real-time or near-real-time feedback about data flow paths, transformations, and issues, enabling administrators to quickly assess data flows without manual investigation. The feedback loop ensures accurate and up-to-date information is always available for decision-making.

Inventive Principle:
Principle #23Feedback

Data Source

PatentUS10958533B2Tracking data flow in distributed computing systems
Publication Date: 2021.03.23 MORGAN STANLEY SERVICES GROUP INC
  • US10958533B2 patent drawing
  • US10958533B2 patent drawing
  • US10958533B2 patent drawing

AI summary

A distributed computing environment comprises a plurality of distributed computer systems that execute a plurality of applications. At least one of the distributed computer systems executes a data flow software tool that identifies potential data flows between the applications and generates a GUI that shows at least one upstream application and/or at least one downstream application for a subject application. The data flow software tool receives, via the GUI, from the user, a first input for the at least one upstream application and/or a second input for the at least one downstream application. The first input comprises a verification that the at least one upstream application provides the incoming data flow to the subject application and the second input comprises a verification that the at least one downstream application receives the outgoing data flow from the subject application.