Voice-Controlled Image Forming System with Pre-Stored Settings

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional image forming systems with interactive agent functions require users to utter multiple settings for job execution, reducing usability.

Innovation Solution

An image forming system that uses a microphone to receive voice inputs, associates image formation settings with identification information, and executes image formation based on acquired settings, allowing users to set and execute jobs using voice commands more efficiently.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If the interactive agent function requires users to utter multiple settings for job execution, then the system can accurately capture all necessary job parameters, but the usability and operation time increase significantly

Engineering Contradiction:
Improvejob setting accuracyVSAvoidusability
Core Design Contradiction:
Measurement precisionVSEase of operation

Solution Approach 1:

The system performs preliminary action by automatically acquiring and storing image formation settings in a storage unit before the user needs them. When a user provides a simple voice input containing identification information, the system retrieves the pre-stored settings associated with that identification, eliminating the need for users to utter multiple setting parameters. This resolves the contradiction by preparing data in advance so that accurate job settings can be obtained through minimal user input.

Inventive Principle:
Principle #10Preliminary action

2Reliability

If the system requires users to provide multiple voice inputs for different settings, then complete job configuration is achieved, but the time required for job execution increases

Engineering Contradiction:
Improvejob configuration completenessVSAvoidjob setup time
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

The system merges multiple settings into a single retrieval operation. Instead of requiring separate voice inputs for each setting parameter, the system combines all image formation settings associated with an identification information into one unified data structure. When the user provides a single voice input containing the identification, the system retrieves and applies all necessary settings simultaneously, ensuring complete job configuration while minimizing the time required.

Inventive Principle:
Principle #5Merging (Combining)

3Ease of operation

If the system stores and manages multiple image formation settings with identification information, then user convenience is improved, but the device complexity increases

Engineering Contradiction:
Improveuser convenienceVSAvoidsystem structure
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The system applies universality by using a single storage unit that can hold multiple types of image formation settings (such as copy settings, FAX settings, print settings) associated with different identification information. This universal storage structure allows the system to manage diverse settings through a unified mechanism, improving user convenience while controlling device complexity through standardized data organization rather than separate specialized storage for each setting type.

Inventive Principle:
Principle #6Universality (Multi-functionality)

Data Source

PatentUS11140284B2Image forming system equipped with interactive agent function, method of controlling same, and storage medium
Publication Date: 2021.10.05 CANON KK
  • US11140284B2 patent drawing
  • US11140284B2 patent drawing
  • US11140284B2 patent drawing

AI summary

An image forming system capable of improving the usability of an interactive agent function. The image forming system receives voice input thereto as an instruction related to execution of a job. The image forming system executes a job based on settings indicated by voice input thereto, and in a case where a specific word is included in the input voice, the image forming system executes the job based on a plurality of types of settings registered in advance in association with the specific word.