Gesture-Based Audio Control for Sound Phrase Manipulation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Users face difficulties in controlling and managing the complex interfaces of electronic devices for creative sound recording and playback, which limits their ability to effectively direct and shape sound recordings and playback, especially in devices with advanced capabilities.

Innovation Solution

A system comprising a processor and memory configured to determine sound characteristics based on user inputs for recording duration and direction, allowing users to control and manipulate sound phrases through a user-friendly interface, enabling the creation and playback of sound phrases with specified characteristics.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If advanced recording devices are used to provide greater capability for creative sound recording and playback, then the capability and functionality of the device is improved, but the device complexity and interface complexity increase, making it harder for average users to control and manage the device

Engineering Contradiction:
Improvecapability for creative sound recording and playbackVSAvoidinterface complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent introduces an intermediary system that translates simple user inputs (taps, slides, draws) into complex audio processing operations. This mediator layer abstracts away the complexity of the underlying audio equipment, allowing users to control sophisticated recording and playback capabilities through intuitive gestures without needing to understand the technical complexity of the device.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent replaces traditional mechanical/physical control interfaces (buttons, knobs, sliders) with gesture-based interaction. Users control audio parameters through movements in the air or on a screen, substituting physical manipulation with motion-based input that is both simpler to operate and more expressive for creative control.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Adaptability or versatility

If complex interfaces are provided to manage recording and playback capabilities, then the functionality and control options are improved, but the ease of operation deteriorates, requiring considerable skill beyond average user capability

Engineering Contradiction:
Improvecontrol options for recording and playbackVSAvoidease of operation
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The patent segments the complex audio control interface into discrete, intuitive gesture categories (tapping for duration, sliding for direction, drawing for spatial mapping). Each gesture type handles a specific control function, breaking down the complex interface into manageable, learnable units that are easy to operate while maintaining full functionality.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

Instead of requiring users to navigate complex menus and adjust multiple parameters to achieve desired audio characteristics, the system inverts the approach by directly mapping simple gestures to audio parameter changes. The interface goes from complex-to-simple rather than simple-to-complex, making operation intuitive and immediate.

Inventive Principle:
Principle #13The other way round (Inversion)

Data Source

PatentEP2660815B1Methods and apparatus for audio processing
Publication Date: 2019.10.23 NOKIA TECHNOLOGIES OY
  • EP2660815B1 patent drawingFigure 1
  • EP2660815B1 patent drawingFigure 2
  • EP2660815B1 patent drawingFigure 3

AI summary

Systems and techniques for determining characteristics to be exhibited by a sound phrase. A user draws traces on an input device indicating characteristics, such as duration and direction, of sounds such as sounds to be captured by a microphone array. In response to the user inputs, signals from the microphone array are processed to produce a signal exhibiting the characteristics. The signal is stored to create a sound phrase, and the sound phrase may later be played. Additional inputs may be received specifying a direction from which the sound phrase is to be played, or playback may come from a default direction. Further inputs may be received during playback to control characteristics of the playback. In addition to specifying characteristics to be imparted to recorded or stored phrases, user inputs may specify characteristics for generated sounds or may specify characteristics to be exhibited by sounds being played.