Voice Profile Synthesis With Secure Authentication and Monetization

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional systems face challenges in accurately reproducing a user's natural voice characteristics, preventing unauthorized use or forgery of generated voice data, and lack mechanisms for users to monetize their own voices, especially when they temporarily or permanently lose the ability to speak.

Innovation Solution

A system that records a user's voice, generates a voice profile, sets parameters for voice synthesis, and assigns a unique identifier to the audio data, utilizing a speech recognition engine for analysis and an identifier generation algorithm to enhance security and data authenticity, enabling accurate reproduction and monetization of synthesized voices.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If conventional voice synthesis systems are used, then basic voice generation is possible, but accurate reproduction of natural voice characteristics cannot be achieved

Engineering Contradiction:
Improvevoice characteristics reproduction accuracyVSAvoidnaturalness of synthesized voice
Core Design Contradiction:
Measurement precisionVSReliability

Solution Approach 1:

The system extracts multiple voice parameters (pitch, tone, rhythm, timbre) from the user's voice sample and uses these parameters to configure the text-to-speech synthesis engine, enabling accurate reproduction of the user's natural voice characteristics while maintaining reliability

Inventive Principle:
Principle #35Parameter changes

Solution Approach 2:

The system creates a digital copy of the user's voice characteristics by analyzing and storing voice profile data, which is then used to generate synthesized speech that closely replicates the original voice without requiring the user to speak in real-time

Inventive Principle:
Principle #26Copying

2Adaptability or versatility

If voice data is generated and stored, then voice synthesis capability is provided, but unauthorized use or forgery of the voice data cannot be prevented

Engineering Contradiction:
Improvevoice utilization capabilityVSAvoidsecurity of voice data
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The system applies cryptographic hashing and digital signature algorithms to the voice data before storage, creating authentication information that prevents unauthorized use or forgery. This preliminary security measure ensures that only authorized applications can access and use the synthesized voice

Inventive Principle:
Principle #9Preliminary anti-action

Solution Approach 2:

The system introduces an authentication layer that acts as an intermediary between the stored voice data and any application seeking to use it. This layer verifies permissions and prevents direct access, thereby securing the voice data while maintaining versatility

Inventive Principle:
Principle #24Intermediary (Mediator)

3Ease of operation

If voice synthesis system is implemented, then users can generate synthetic voice, but mechanisms for users to monetize their own voices are lacking

Engineering Contradiction:
Improveuser accessibility to voice synthesisVSAvoidmonetization capability
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The system provides multiple functions within a single platform: voice recording, analysis, synthesis generation, secure storage, and monetization capabilities. Users can easily access voice synthesis and simultaneously utilize the platform to commercialize their synthesized voices through various applications

Inventive Principle:
Principle #6Universality (Multi-functionality)

4Ease of manufacture

If simple voice recording is performed, then basic audio data is obtained, but accurate voice profile analysis for synthesis cannot be conducted

Engineering Contradiction:
Improvesimplicity of voice recording processVSAvoidvoice profile analysis accuracy
Core Design Contradiction:
Ease of manufactureVSMeasurement precision

Solution Approach 1:

The system performs preliminary voice analysis by extracting multiple parameters (pitch, tone, rhythm, timbre) from the recorded voice sample before synthesis. This preliminary action ensures that even simple recordings can be accurately analyzed to create detailed voice profiles for high-quality synthesis

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS20260051314A1system
Publication Date: 2026.02.19 SOFTBANK GROUP CORP
  • US20260051314A1 patent drawing
  • US20260051314A1 patent drawing
  • US20260051314A1 patent drawing

AI summary

A system includes a processor that is configured to record a user's voice, transmit the recorded voice to a server, analyze the audio data at the server to generate a voice profile, perform transcription based on the voice profile, set parameters for voice synthesis, assign a unique identifier to the audio data, generate synthetic voice, and enable the user to download the generated voice data.