Centralized AutoML Platform for Disparate Real-Time Datasets

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing systems struggle to efficiently analyze and forecast using large, disparate datasets generated by entities, as they often require substantial computing power, specific dataset formatting, and technical expertise, limiting their ability to handle real-time updates and inter-relations within these datasets.

Innovation Solution

A centralized platform that automatically imports and trains machine learning models on diverse datasets, using APIs to integrate data sources, train models based on heuristics and templates, and provide real-time insights through simplified user interfaces.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If electronic spreadsheet applications are used to analyze datasets, then ease of operation is improved, but the system cannot cope with large quantities of disparate datasets and real-time updates

Engineering Contradiction:
Improveease of operationVSAvoidadaptability
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

An automated machine learning platform serves as an intermediary between electronic spreadsheet applications and complex machine learning techniques. The platform automatically ingests disparate datasets, trains multiple ML models, and delivers results to spreadsheet applications, enabling users to leverage advanced ML capabilities without directly managing the complexity of data ingestion, model training, and evaluation.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Adaptability or versatility

If automated machine learning platforms are used to handle disparate datasets, then adaptability is improved, but computing resources and system complexity increase

Engineering Contradiction:
ImproveadaptabilityVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The automated machine learning platform performs self-service by automatically ingesting datasets from multiple sources, selecting appropriate models, training models, evaluating performance, and delivering results without requiring manual intervention for each step. This automation handles the complexity internally while presenting a simplified interface to users.

Inventive Principle:
Principle #25Self-service

3Device complexity

If traditional systems are used to analyze datasets, then system complexity is reduced, but the ability to rapidly ingest and analyze fresh data is limited

Engineering Contradiction:
Improvesystem complexityVSAvoidproductivity
Core Design Contradiction:
Device complexityVSProductivity

Solution Approach 1:

The system performs preliminary actions by pre-configuring multiple machine learning models and training pipelines in advance. When new datasets arrive, the pre-configured system can immediately begin processing without requiring manual setup or configuration, enabling rapid ingestion and analysis of fresh data while maintaining manageable system complexity through standardized workflows.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS20260094009A1Centralized platform for enhanced automated machine learning using disparate datasets
Publication Date: 2026.04.02 AMAZON TECH INC
  • US20260094009A1 patent drawing
  • US20260094009A1 patent drawing
  • US20260094009A1 patent drawing

AI summary

Systems and techniques are disclosed for a centralized platform for enhanced automated machine learning using disparate datasets. An example method includes receiving user specification of one or more data sources to be integrated with the system, the data sources storing datasets to be utilized to train one or more machine learning models by the system, and the datasets reflecting user interaction data. A dataset is imported from the data source, and machine learning models are automatically trained based a particular machine learning model recipe of a plurality of machine learning model recipes. A first trained machine learning model is implemented, with the system being configured to respond to queries based on the implemented machine learning model, and with the responses including personalized recommendations.