Text-to-Speech Interface with Ranked Phrase Library

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing techniques for generating synthesized speech outputs on electronic devices are cumbersome and inefficient, requiring complex user interfaces that consume time and energy, particularly in battery-operated devices.

Innovation Solution

The implementation of faster and more efficient methods and interfaces that provide a candidate text library with ranked and categorized phrases, allowing for quick selection and conversion to synthesized speech, along with intuitive controls for text input and output management, using a personalized voice model to enhance speech output effectiveness.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If complex user interfaces with multiple inputs are used to control text entry, then text input capability is achieved, but user time and device energy are increased

Engineering Contradiction:
Improvetext input efficiencyVSAvoidtime to generate speech output
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The system pre-generates and stores a library of candidate text phrases that are likely to be useful for speech output. When the user activates the text-to-speech function, these pre-prepared phrases are immediately available for selection and conversion, eliminating the time-consuming process of manual text entry while maintaining comprehensive text input capability

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system creates visual copies (display representations) of candidate text phrases that the user can select. Instead of requiring the user to type or complexly input text, the system presents copy-ready text options that can be quickly selected and converted to speech, dramatically reducing the time and effort required for text input

Inventive Principle:
Principle #26Copying

2Ease of operation

If complex user interfaces with multiple inputs are used to control text entry, then text input capability is achieved, but device energy consumption is increased

Engineering Contradiction:
Improvetext input efficiencyVSAvoiddevice energy consumption
Core Design Contradiction:
Ease of operationVSUse of energy by moving object

Solution Approach 1:

The candidate text library is generated and prepared in advance, storing frequently used or contextually relevant phrases. When text-to-speech is activated, the system immediately accesses this pre-prepared library rather than generating text in real-time through complex user input processes, significantly reducing the computational energy required during active use

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system provides self-service by automatically managing the candidate text library and presenting relevant phrases to the user. This eliminates the need for energy-intensive real-time text generation and complex interface processing, as the system serves itself by maintaining and managing the phrase library autonomously

Inventive Principle:
Principle #25Self-service

3Reliability

If personalized voice models are used to enhance speech output, then speech effectiveness is improved, but processing complexity is increased

Engineering Contradiction:
Improvespeech output effectivenessVSAvoidprocessing complexity
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The system applies personalized voice modeling selectively to specific candidate phrases rather than processing all possible text inputs. By focusing the personalized processing only on the limited set of pre-generated candidate phrases, the system achieves high speech effectiveness while keeping overall processing complexity manageable through localized application of the personalization model

Inventive Principle:
Principle #3Local quality

Data Source

PatentUS20240386877A1Techniques and user interfaces for generating synthesized speech
Publication Date: 2024.11.21 APPLE INC
  • US20240386877A1 patent drawing
  • US20240386877A1 patent drawing
  • US20240386877A1 patent drawing

AI summary

The present disclosure generally relates to techniques and interfaces for generating synthesized speech outputs. For example, a user interface for a text-to-speech service can include ranked and/or categorized phrases, which can be selected to enter as text. A synthesized speech output is then generated to deliver any entered text, for example, using a personalized voice model.