Sound Processing Unit for Text-Based Observation Site Recording

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional information recording systems struggle to efficiently record and retrieve detailed information about observation sites, especially in environments where manual recording is difficult or unsafe, leading to potential omissions or errors in documenting situations and events.

Innovation Solution

An information recording system that includes a sound acquisition unit, sound processing unit, recording unit, keyword reception unit, search unit, and display unit, which converts sound information into text and associates it with object information and time point information, allowing for efficient keyword searches and display of relevant information.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If sound information is converted into text information and recorded in association with object information, then the accuracy and completeness of observation site recording is improved, but the device complexity and processing time increase

Engineering Contradiction:
Improverecording accuracyVSAvoidsystem complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent introduces a sound processing unit as an intermediary component that converts sound information into text information. This mediator handles the complex speech-to-text conversion process, isolating the complexity from the main recording system while enabling accurate text-based recording of observation sites and events.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The recording system is divided into distinct functional modules: sound acquisition unit, sound processing unit, recording unit, search unit, and display unit. Each module handles a specific aspect of the recording process, allowing the system to manage complexity through modular design while maintaining high recording accuracy.

Inventive Principle:
Principle #1Segmentation

2Productivity

If sound information is converted into text information and recorded in association with object information and time point information, then the efficiency of information retrieval is improved, but the loss of time for processing increases

Engineering Contradiction:
Improveretrieval efficiencyVSAvoidprocessing time
Core Design Contradiction:
ProductivityVSLoss of time

Solution Approach 1:

The system performs sound-to-text conversion and associates text information with object information and time point information in advance, during the recording phase. This preliminary processing enables rapid retrieval later, as the information is already structured and indexed with temporal and contextual markers, reducing retrieval time despite the initial processing overhead.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system creates text copies of sound information that can be independently searched and retrieved without needing to process the original audio data. These text copies, enriched with object information and time point metadata, serve as efficient proxies for rapid information access.

Inventive Principle:
Principle #26Copying

3Device complexity

If only object information is recorded, then the simplicity of the recording system is maintained, but the ability to record detailed situation information is insufficient

Engineering Contradiction:
Improvesystem simplicityVSAvoidsituation detail
Core Design Contradiction:
Device complexityVSLoss of information

Solution Approach 1:

The recording system is enhanced to perform multiple functions: it records not only object information but also sound information, converts sound to text, and associates all three types of information (object, sound, and text) with time point information. This multi-functional approach comprehensively captures observation site details without requiring separate specialized systems.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The patent merges object information, sound information, and time point information into a unified recording structure. By combining these different information types and associating them through the recording unit, the system preserves detailed situation information that would be lost if only object information were recorded.

Inventive Principle:
Principle #5Merging (Combining)

Applied Scientific Principles

This section explains which scientific principles are used to turn an abstract innovation direction into a practical engineering solution.

Function Achieved in This Case

The system enables accurate and efficient recording and retrieval of observation site information, reducing omissions and errors by converting sound into text and associating it with object and time point data, facilitating easy access to specific events or situations.

Implementation Method 1

The sound processing unit converts the sound information acquired by the sound acquisition unit into text information

Methodology Applied
Scientific EffectSpeech recognition:

Data Source

PatentUS11036775B2Information recording system and information recording method
Publication Date: 2021.06.15 OLYMPUS CORPORATION(JP)
  • US11036775B2 patent drawing
  • US11036775B2 patent drawing
  • US11036775B2 patent drawing

AI summary

In an information recording system, a sound processing unit generates a conversion candidate word in a process of converting sound information into text information. A recording unit records the text information and the conversion candidate word on a recording medium such that the text information and the conversion candidate word are associated with each other. A search unit performs a search based on a keyword and extracts a word matching the keyword from words within the text information and the conversion candidate word. A reading unit reads the text information including the word matching the keyword from the recording medium. A display unit displays the text information such that a part corresponding to the word matching the keyword and a part other than the corresponding part are able to be distinguished.