Centralized Voice Model Server for Audio Content Redundancy

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Disparate computing resources face challenges in efficiently processing and consistently providing audio-based content items due to lack of access to the same voice models or outdated models, leading to redundant processing and inefficient resource utilization.

Innovation Solution

A data processing system that modifies computing program output by parsing voice-based instructions, selecting or reusing parameterized content items, and routing them through a dialog data structure, allowing for automated native content provision and session management to reduce redundant processing and improve resource utilization.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If disparate computing resources process audio-based content items independently using their own voice models, then each resource can operate autonomously, but processing redundancy increases and resource utilization efficiency decreases

Engineering Contradiction:
Improveautonomous operation capabilityVSAvoidresource utilization efficiency
Core Design Contradiction:
Adaptability or versatilityVSProductivity

Solution Approach 1:

The patent merges voice model processing across disparate computing resources by implementing a centralized voice model server that stores and manages unified voice models. Multiple chatbots query this central server for content items instead of maintaining separate voice models locally, thereby eliminating processing redundancy while preserving autonomous operation through the distributed query architecture.

Inventive Principle:
Principle #5Merging (Combining)

Solution Approach 2:

The centralized voice model server provides universal access to voice models for multiple chatbots and computing resources. A single voice model repository serves multiple purposes: storing voice models, managing content items, and providing unified access points for all chatbots, thereby improving resource utilization across the entire system.

Inventive Principle:
Principle #6Universality (Multi-functionality)

2Ease of operation

If computing resources maintain separate voice models and content item databases, then each resource can operate independently, but redundant processing occurs and processor utilization decreases

Engineering Contradiction:
Improveindependent operation capabilityVSAvoidprocessor utilization
Core Design Contradiction:
Ease of operationVSPower

Solution Approach 1:

The patent introduces a centralized voice model server as an intermediary between disparate computing resources and voice model content. This mediator handles all queries for content items, maintaining independent operation of individual chatbots while centralizing processing to eliminate redundancy and improve overall processor utilization.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Productivity

If multiple chatbots query for the same content items independently, then each chatbot can respond autonomously to user requests, but network traffic increases and bandwidth utilization becomes inefficient

Engineering Contradiction:
Improveresponse capabilityVSAvoidbandwidth utilization efficiency
Core Design Contradiction:
ProductivityVSLoss of energy

Solution Approach 1:

The system implements preliminary action by having chatbots query the centralized voice model server for content items before actual user interaction requires them. The centralized server prepares and caches content items in advance, so when multiple chatbots need the same content, it is already available and can be distributed efficiently, reducing redundant network traffic.

Inventive Principle:
Principle #10Preliminary action

4Adaptability or versatility

If computing resources use outdated or unsynchronized voice models, then local processing can continue without external dependencies, but audio content accuracy and consistency deteriorate

Engineering Contradiction:
Improvelocal processing capabilityVSAvoidaudio content consistency
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The centralized voice model server implements feedback mechanisms where chatbots report their content item needs and the server responds by providing updated content items. This feedback loop ensures that all computing resources receive synchronized, accurate voice models while maintaining the ability to process locally, thereby improving audio content consistency without sacrificing local processing capability.

Inventive Principle:
Principle #23Feedback

Data Source

PatentEP3430616B1Modification of audio-based computer program output
Publication Date: 2024.08.07 GOOGLE LLC
  • EP3430616B1 patent drawingFigure 1
  • EP3430616B1 patent drawingFigure 2
  • EP3430616B1 patent drawingFigure 3

AI summary

Modifying computer program output in a voice or non-text input activated environment is provided. A system can receive audio signals detected by a microphone of a device. The system can parse the audio signal to identify a computer program to invoke. The computer program can identify a dialog data structure. The system can modify the identified dialog data structure to include a content item. The system can provide the modified dialog data structure to a computing device for presentation.