Voice Driven Dynamic Menu for Video Audio Modification

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current social networking platforms lack effective methods for enhancing communication and content sharing, particularly in modifying audio tracks within videos, which limits user engagement and interaction.

Innovation Solution

A messaging system that dynamically analyzes audio portions of videos to determine the presence of a voice signal, presenting different menus for modifying the audio track, allowing users to enhance or alter the voice signal, and storing the modified audio with the video track.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If dynamic menu options are provided for voice signal modification, then user engagement and content sharing experiences are improved, but device complexity and processing requirements increase

Engineering Contradiction:
Improveuser engagementVSAvoidprocessing requirements
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The system performs preliminary voice signal detection and analysis before presenting menu options to the user. By pre-processing the audio track to identify voice signals and determine appropriate modification options, the system reduces the complexity of real-time processing during user interaction while maintaining ease of operation.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent introduces an intermediary processing layer that analyzes audio tracks and generates contextual menu options. This intermediary system acts as a mediator between the raw audio data and the user interface, automatically detecting voice signals and presenting relevant modification options, thereby simplifying the user experience while managing processing complexity through automated intermediation.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Adaptability or versatility

If audio track modification capabilities are added to enhance communication, then functionality and user engagement improve, but device complexity increases

Engineering Contradiction:
ImprovefunctionalityVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent implements a universal audio processing framework that can detect voice signals, generate contextual menus, and apply various modifications (frequency adjustment, pitch correction, etc.) through a single integrated system. This multi-functional approach allows the same core infrastructure to handle diverse audio modification tasks, enhancing versatility while avoiding the need for separate complex systems for each function.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The system employs dynamic menu generation that adapts to the detected audio content. Rather than providing static modification options, the system dynamically creates context-relevant menus based on the analyzed voice signal characteristics, allowing the interface to adapt its functionality to the specific audio being processed, thereby enhancing versatility without requiring all possible modification capabilities to be simultaneously available.

Inventive Principle:
Principle #15Dynamics

Data Source

PatentUS11934636B2Voice driven dynamic menus
Publication Date: 2024.03.19 SNAP INC
  • US11934636B2 patent drawing
  • US11934636B2 patent drawing
  • US11934636B2 patent drawing

AI summary

Disclosed are systems, methods, and computer-readable storage media to provide voice driven dynamic menus. One aspect disclosed is a method including receiving, by an electronic device, video data and audio data, displaying, by the electronic device, a video window, determining, by the electronic device, whether the audio data includes a voice signal, displaying, by the electronic device, a first menu in the video window in response to the audio data including a voice signal, displaying, by the electronic device, a second menu in the video window in response to a voice signal being absent from the audio data, receiving, by the electronic device, input from the displayed menu, and writing, by the electronic device, to an output device based on the received input.