Voice-Controlled Image Forming Apparatus Template Input
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing image forming systems cannot conveniently input and print voice-instructed character strings into templates with text input fields, nor can they search for and utilize image data as intended by the user through pronunciation for image formation.
Innovation Solution
An information processing apparatus with a communication interface and a control device that recognizes voice input from a smart speaker, specifies data from the recognized voice content, adds it to a designated template, and transmits a command for image formation to an image forming apparatus, enabling voice-instructed input and the use of searched image data for printing.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of operation
If voice input is used to input character strings into templates, then ease of operation is improved, but the system cannot recognize and process the voice content to add data to templates
Solution Approach 1:
The patent introduces a smart speaker as an intermediary device between the user and the image forming apparatus. The smart speaker captures voice input and transmits it to the image forming apparatus, which then processes the voice content through speech-to-text conversion. This intermediary approach enables voice-based template data input without requiring direct voice processing integration in the traditional printing system.
Solution Approach 2:
The patent replaces traditional mechanical input methods (keyboard typing, manual data entry) with voice-based input. The image forming apparatus converts voice signals into text data that can be added to templates, substituting the mechanical interaction with an acoustic field-based interaction. This allows users to input characters and data into templates through speech rather than physical keyboard input.
2Productivity
If traditional input methods are used, then system complexity is reduced, but productivity and user convenience are worsened
Solution Approach 1:
The patent makes the image forming apparatus multi-functional by integrating speech-to-text conversion capabilities and template management functions. The apparatus can now handle both traditional printing tasks and voice-based data input, allowing a single device to perform multiple functions. This universality improves productivity without requiring separate dedicated devices for voice processing and printing.
Solution Approach 2:
The patent combines the smart speaker's voice capture capability with the image forming apparatus's printing and template processing functions. By merging these previously separate functions into an integrated workflow, the system enables voice input to directly generate printable content. The control device merges voice recognition results with template data structures, creating a unified input-processing-output system that improves efficiency.
Data Source
AI summary
An information processing apparatus includes: a communication interface; and a control device configured to: recognize a content of voice input by utterance of a user of an image forming apparatus from a smart speaker connected via the communication interface configured to input and output voice; and in a case the recognized content of voice includes designating a template and adding data to a template, specify the data from the recognized content of voice, add the specified data to the designated template, and transmit a command for image formation to the image forming apparatus.


