Voice-Controlled Image Forming Apparatus Template Input

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing image forming systems cannot conveniently input and print voice-instructed character strings into templates with text input fields, nor can they search for and utilize image data as intended by the user through pronunciation for image formation.

Innovation Solution

An information processing apparatus with a communication interface and a control device that recognizes voice input from a smart speaker, specifies data from the recognized voice content, adds it to a designated template, and transmits a command for image formation to an image forming apparatus, enabling voice-instructed input and the use of searched image data for printing.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If voice input is used to input character strings into templates, then ease of operation is improved, but the system cannot recognize and process the voice content to add data to templates

Engineering Contradiction:
Improveease of inputVSAvoidvoice processing capability
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The patent introduces a smart speaker as an intermediary device between the user and the image forming apparatus. The smart speaker captures voice input and transmits it to the image forming apparatus, which then processes the voice content through speech-to-text conversion. This intermediary approach enables voice-based template data input without requiring direct voice processing integration in the traditional printing system.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent replaces traditional mechanical input methods (keyboard typing, manual data entry) with voice-based input. The image forming apparatus converts voice signals into text data that can be added to templates, substituting the mechanical interaction with an acoustic field-based interaction. This allows users to input characters and data into templates through speech rather than physical keyboard input.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Productivity

If traditional input methods are used, then system complexity is reduced, but productivity and user convenience are worsened

Engineering Contradiction:
Improveinput efficiencyVSAvoidsystem complexity
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent makes the image forming apparatus multi-functional by integrating speech-to-text conversion capabilities and template management functions. The apparatus can now handle both traditional printing tasks and voice-based data input, allowing a single device to perform multiple functions. This universality improves productivity without requiring separate dedicated devices for voice processing and printing.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Solution Approach 2:

The patent combines the smart speaker's voice capture capability with the image forming apparatus's printing and template processing functions. By merging these previously separate functions into an integrated workflow, the system enables voice input to directly generate printable content. The control device merges voice recognition results with template data structures, creating a unified input-processing-output system that improves efficiency.

Inventive Principle:
Principle #5Merging (Combining)

Data Source

PatentUS11474782B2Information processing apparatus, information processing method and non-transitory computer-readable medium
Publication Date: 2022.10.18 BROTHER KOGYO KK
  • US11474782B2 patent drawing
  • US11474782B2 patent drawing
  • US11474782B2 patent drawing

AI summary

An information processing apparatus includes: a communication interface; and a control device configured to: recognize a content of voice input by utterance of a user of an image forming apparatus from a smart speaker connected via the communication interface configured to input and output voice; and in a case the recognized content of voice includes designating a template and adding data to a template, specify the data from the recognized content of voice, add the specified data to the designated template, and transmit a command for image formation to the image forming apparatus.