Community Audio Narration System Using Segmented Human Recordings

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Users of digital content face limitations in enjoying audio versions of text-based works while performing tasks like driving or running, as not all content has audio versions, and authors lack resources for human narrators, leading to subpar computer-synthesized audio.

Innovation Solution

A community-based audio narration system that leverages a network of human readers to contribute and combine audio recordings, using a computing environment with server modules for content presentation, audio collection, analysis, integration, and distribution, to create natural-sounding audio readings without relying on professional narrators.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of manufacture

If text-to-speech technology is used to convert digital content to synthesized speech, then audio versions of text-based works can be produced without human narrators, but the synthesized speech sounds unnatural and stilted

Engineering Contradiction:
Improveability to produce audio versionsVSAvoidnaturalness of speech
Core Design Contradiction:
Ease of manufactureVSReliability

Solution Approach 1:

The patent segments the audio production process into multiple independent contributions from different community members. Each user records specific sections or chapters rather than requiring one person to record the entire work, allowing the system to aggregate many small, natural-sounding segments into a complete audio version.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent implements a self-service model where the community itself produces the audio narrations without requiring professional narrators or expensive text-to-speech systems. Users voluntarily contribute their voice recordings to create audio versions of text-based works, making the system self-sustaining and cost-effective.

Inventive Principle:
Principle #25Self-service

2Reliability

If professional human narrators are hired to produce audio versions, then natural-sounding narrations can be achieved, but the cost and resource requirements increase significantly

Engineering Contradiction:
Improvenaturalness of speechVSAvoidcost and resource requirements
Core Design Contradiction:
ReliabilityVSEase of manufacture

Solution Approach 1:

The system replaces expensive professional narrators with a self-service model where any user can contribute audio recordings. This eliminates the need for hiring professionals while maintaining natural speech quality, as real people are recording the content rather than relying on synthesized speech.

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent creates multiple copies of audio recordings from different users for the same text sections. These copies are then aggregated and integrated into a unified audio version, allowing the system to leverage many contributors without needing each to be a professional narrator.

Inventive Principle:
Principle #26Copying

3Ease of manufacture

If a community-based system is used to collect audio recordings from multiple users, then natural-sounding narrations can be produced cost-effectively, but the system complexity increases

Engineering Contradiction:
Improvecost-effectivenessVSAvoidsystem complexity
Core Design Contradiction:
Ease of manufactureVSDevice complexity

Solution Approach 1:

The patent introduces an intermediary platform that manages the complex coordination between multiple users, text sections, and audio recordings. This intermediary system handles user registration, text segmentation, recording submission, quality verification, and audio integration, shielding end users from the underlying complexity.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The system implements feedback mechanisms where users can rate and review audio contributions from other users. This feedback loop allows the community to self-regulate quality and provides guidance for improving future recordings, reducing the need for complex manual quality control.

Inventive Principle:
Principle #23Feedback

Data Source

PatentUS9002703B1Community audio narration generation
Publication Date: 2015.04.07 AMAZON TECH INC
  • US9002703B1 patent drawing
  • US9002703B1 patent drawing
  • US9002703B1 patent drawing

AI summary

The community-based generation of audio narrations for a text-based work leverages collaboration of a community of people to provide human-voiced audio readings. During the community-based generation, a collection of audio recordings for the text-based work may be collected from multiple human readers in a community. An audio recording for each section in the text-based work may be selected from the collection of audio recordings. The selected audio recordings may be then combined to produce an audio reading of at least a portion of the text-based work.