Audio Guidance Generation Device for Sports Broadcasts

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing methods for generating audio guidance during sports broadcasts are limited, as they do not effectively convey the situation of a competition in conjunction with video and often rely on human commentators, which can be costly and delayed, especially in large-scale events.

Innovation Solution

An audio guidance generation device that accumulates and analyzes competition data to generate explanatory text and speech, using a message management unit, explanation generation unit, and speech synthesis unit to provide real-time audio guidance, detecting unconveyed information and updating explanations based on new data, and incorporating phoneme language features and acoustic models for intonation.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If live commentary is provided by human announcers at game venues, then the quality and understanding of sports programs is improved, but the cost increases significantly

Engineering Contradiction:
Improvequality of sports programVSAvoidcost
Core Design Contradiction:
ReliabilityVSQuantity of substance

Solution Approach 1:

The patent uses automatic text generation to create commentary that copies and replaces human announcer output. The system generates explanatory text from competition data using templates, effectively copying the informational function of human commentators without the associated costs.

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The patent replaces the mechanical system of human announcers with an automated text generation system. This substitution uses computer-based text generation mechanisms to perform the commentary function that previously required human vocal and cognitive systems.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Reliability

If live commentary is provided by human announcers, then the understanding of competition situation is improved, but the response time is delayed due to processing and subtitle generation

Engineering Contradiction:
Improveunderstanding of competition situationVSAvoidresponse time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The patent performs preliminary action by pre-defining explanation templates and structures before competition events occur. These templates are prepared in advance with all possible explanation patterns, allowing immediate text generation when competition data is received, eliminating the time needed for real-time human processing and subtitle creation.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent replaces the slow mechanical processes of human speech production and subtitle generation with instant computer-based text generation. The automated system processes competition data and generates explanatory text immediately, eliminating the sequential delays inherent in human commentary workflows.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

3Quantity of substance

If automatic text generation is used for commentary, then the cost is reduced, but the text does not effectively convey the situation in conjunction with video

Engineering Contradiction:
ImprovecostVSAvoideffectiveness of situation conveyance
Core Design Contradiction:
Quantity of substanceVSReliability

Solution Approach 1:

The patent segments the commentary generation process into distinct functional components: competition data reception, unconveyed information detection, template selection, and text generation. This segmentation allows each component to specialize in its function, ensuring that the final text effectively conveys competition situations while maintaining cost efficiency.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent implements feedback by detecting what information has already been conveyed and using that detection to guide subsequent text generation. The system continuously monitors the competition situation and adjusts its explanations based on what has already been communicated, ensuring effectiveness without redundancy.

Inventive Principle:
Principle #23Feedback

4Reliability

If commentary broadcast is produced with multiple people, then the quality is improved, but the complexity and resource requirements increase

Engineering Contradiction:
Improvequality of commentary broadcastVSAvoidnumber of people required
Core Design Contradiction:
ReliabilityVSDevice complexity

Solution Approach 1:

The patent creates a universal text generation system that performs multiple functions previously requiring different people: detecting unconveyed information, selecting appropriate explanations, and generating commentary text. This single automated system replaces multiple human roles, reducing complexity while maintaining quality.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The patent implements self-service by enabling the commentary system to automatically detect what information needs to be conveyed and generate appropriate explanations without human intervention. The system serves itself by autonomously managing the entire commentary generation process from data reception to text production.

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS11404041B2Audio guidance generation device, audio guidance generation method, and broadcasting system
Publication Date: 2022.08.02 NIPPON HOSO KYOKAI
  • US11404041B2 patent drawing
  • US11404041B2 patent drawing
  • US11404041B2 patent drawing

AI summary

A message management unit receives and accumulates a message, wherein the message is distributed for every update, is the message data representing a latest situation of a competition, an explanation generation unit generates an explanatory text for conveying unconveyed information detected from the message, based on conveyed information, a speech synthesis unit outputs a speech converted from the explanatory text, wherein the explanation generation unit stores the unconveyed information for the explanatory text as the conveyed information, stands by until completion of completion of the speech, and initiates a procedure for generating a new explanatory text based on updated unconveyed information.