Image forming apparatus

The image forming apparatus facilitates verbal communication by integrating voice input and storage, addressing impersonal communication in remote work environments.

JP2025165508APending Publication Date: 2025-11-05KYOCERA DOCUMENT SOLUTIONS INC
View PDF 4 Cites 0 Cited by

Patent Information

Application Number
JP2024069584
Authority / Receiving Office
JP · JP
Patent Type
Applications
Current Assignee / Owner
Filing Date
2024-04-23
Publication Date
2025-11-05

AI Technical Summary

Technical Problem

In remote work environments, direct verbal communication is decreasing, leading to impersonal communication between users exchanging printed materials.

Method used

An image forming apparatus that integrates voice input and storage capabilities, associating unique identification information with audio data to enable verbal communication alongside paper media.

Benefits of technology

Enables oral communication of information expressed in spoken language, allowing for appropriate and smooth communication even in remote work settings.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 2025165508000001_ABST
    Figure 2025165508000001_ABST
Patent Text Reader

Abstract

To provide an image forming apparatus that can achieve appropriate and smooth communication by voice together with a paper medium even in a remote work environment.SOLUTION: A control unit 10 generates unique identification information according to a predetermined rule along with creation of a print by an image forming unit 12, inputs the voice of a user to an input unit 21, and stores the generated identification information and voice data indicating the voice inputted to the input unit 21 in a predetermined storage destination in association with each other.SELECTED DRAWING: Figure 5
Need to check novelty before this filing date? Find Prior Art

Description

[Technical Field]

[0001] The present invention relates to an image forming apparatus, and more particularly to a technique for registering voice data in association with a printed matter. [Background technology]

[0002] Generally, various devices that handle audio data are known. For example, Patent Document 1 discloses a technology in which multiple image forming devices transmit and receive audio data indicating communication sounds via an exchange. Patent Document 2 discloses a digital camera that records captured images in association with audio on a memory card, displays a list of captured images on a display unit while playing the audio, and plays the audio corresponding to an image selected by a user.

[0003] Patent Document 3 discloses a technology that uses audio mapping data to play both audio data acquired from a source other than an information recording medium and video data stored on a recording medium. Patent Document 4 discloses a technology that associates image files and audio files and stores them on a CD-R, and then plays the images and audio on a personal computer. [Prior art documents] [Patent documents]

[0004] [Patent Document 1] Japanese Patent Application Publication No. 2018-148401 [Patent Document 2] Japanese Patent Application Laid-Open No. 2008-085486 [Patent Document 3] Special Publication No. 2007-501485 [Patent Document 4] Japanese Patent Application Laid-Open No. 2003-122607 Summary of the Invention [Problem to be solved by the invention]

[0005] However, with the spread of remote work, opportunities for direct verbal communication are decreasing. For example, in a remote work environment, the user who created the printed material sends it to the remote work location of another user. In such a situation, communication between users tends to be indirect and impersonal.

[0006] The present invention has been made in consideration of the above circumstances, and aims to provide an image forming apparatus that can realize appropriate and smooth communication by voice as well as paper media even in a remote work environment. [Means for solving the problem]

[0007] An image forming apparatus according to one aspect of the present invention comprises an image forming unit that forms an image indicated by image data on recording paper to generate a printed matter, an input unit into which a user's voice is input, and a control unit that, as the image forming unit generates a printed matter, causes the user's voice to be input into the input unit, generates unique identification information in accordance with predetermined rules, and associates the generated identification information with audio data indicating the voice input into the input unit and stores the generated identification information in a predetermined storage location. [Effects of the Invention]

[0008] According to the present invention, a user who creates a printed material can save audio related to the created printed material, and other users who acquire the printed material can use the identification information to identify and acquire the audio related to the created printed material. This enables oral communication of information expressed in spoken language that cannot be obtained by exchanging printed materials alone. Therefore, even in a remote work environment, appropriate and smooth communication can be achieved using audio as well as paper media. [Brief explanation of the drawings]

[0009] [Figure 1] FIG. 2 is a front cross-sectional view showing the structure of the image forming apparatus. [Figure 2]FIG. 1 is a block diagram showing a configuration of an image forming apparatus. [Figure 3] FIG. 10 is a diagram illustrating an example of an authentication screen. [Figure 4] FIG. 10 is a diagram illustrating an example of a job selection screen. [Figure 5] 10 is a flowchart showing a first voice data registration process. [Figure 6] FIG. 10 is a diagram illustrating an example of a print start screen. [Figure 7] FIG. 10 is a diagram illustrating an example of a voice input screen. [Figure 8] FIG. 10 is a diagram illustrating an example of a destination input screen. [Figure 9] 10 is a flowchart showing a first audio data reproduction process. [Figure 10] FIG. 10 is a diagram illustrating an example of an identification information input screen. [Figure 11] FIG. 10 is a diagram illustrating an example of a password input screen. [Figure 12] FIG. 10 is a diagram showing an example of a playback instruction input screen. [Figure 13] 10 is a flowchart showing a second audio data reproduction process. [Figure 14] FIG. 10 is a diagram showing another example of the identification information input screen. DETAILED DESCRIPTION OF THE INVENTION

[0010] An image forming apparatus according to one embodiment of the present invention will be described below with reference to the drawings. Fig. 1 is a front cross-sectional view showing the structure of an image forming apparatus 1 according to one embodiment of the present invention. Fig. 2 is a block diagram showing the configuration of the image forming apparatus 1.

[0011] [Device configuration] 1, image forming apparatus 1 is a multifunction device having multiple functions such as a copy function, a transmission function, a printer function, and a facsimile function. The housing of image forming apparatus 1 houses an image reading unit 11, an image forming unit 12, a fixing unit 13, a paper feeding unit 14, etc.

[0012] The image reading unit 11 is an ADF (Auto Document Feeder) that includes a document transport unit 6 that transports a document placed on a document table, and a scanner that optically reads a document transported by the document transport unit 6 or a document placed on a platen glass 7. The image reading unit 11 reads the image of the document by illuminating the document with a light irradiation unit and receiving reflected light with a CCD (Charge-Coupled Device) sensor, thereby generating image data representing the document image.

[0013] The image forming unit 12 includes a photosensitive drum, a charging device, an exposure device, a developing device, and a transfer device. Based on the image data generated by the image reading unit 11 or the image data input via the communication unit 24, the image forming unit 12 forms an image composed of a toner image on the recording paper P conveyed along the conveying path T by the conveying unit 17, thereby generating a printed matter.

[0014] The fixing unit 13 applies heat and pressure to the recording paper P on which the toner image has been formed by the image forming unit 12, thereby fixing the toner image to the recording paper P. The recording paper P on which the toner image has been fixed is discharged to the discharge tray 8.

[0015] The paper feed unit 14 includes a manual feed tray and multiple paper feed cassettes. The paper feed unit 14 uses a pickup roller to pull out recording paper P stored in one of the multiple paper feed cassettes or recording paper placed on the manual feed tray one sheet at a time and feeds the paper to the transport path T.

[0016] 2, the image forming apparatus 1 includes a control unit 100. The control unit 100 includes a processor, a random access memory (RAM), and a read only memory (ROM). The processor is, for example, a central processing unit (CPU), a micro processing unit (MPU), or an application specific integrated circuit (ASIC).

[0017] The control unit 100 is electrically connected to the document transport unit 6, image reading unit 11, image forming unit 12, fixing unit 13, paper feeding unit 14, display unit 15, operation unit 16, transport unit 17, HDD (Hard Disk Drive) 18, image processing unit 19, image memory 20, input unit 21, output unit 22, facsimile communication unit 23, and communication unit 24, etc.

[0018] The control unit 100 functions as the control unit 10 by the processor executing various computer programs stored in the built-in ROM or HDD 18. The control unit 10 performs overall control of the image forming apparatus 1. More specifically, the control unit 10 controls the operation of each part of the image forming apparatus 1 and communication with external devices connected via a network.

[0019] The control unit 10 may be configured to be operable by a logic circuit, not by operation based on a computer program, or may be configured to be realized by two or more control units.

[0020] The display unit 15 is a display device configured by a liquid crystal display or an organic light-emitting diode (EL) display. The display unit 15 displays various screens for the functions that can be executed by the image forming apparatus 1.

[0021] The operation unit 16 includes a plurality of hard keys including a start key 16A. The operation unit 16 also includes a touch panel 16B that is placed on top of the display unit 15. User instructions regarding each function that can be executed by the image forming apparatus 1 are input to the operation unit 16.

[0022] The conveying unit 17 includes a pair of conveying rollers 17A, a pair of discharge rollers 17B, and a conveying motor. The control unit 10 drives the conveying motor to rotate the pair of conveying rollers 17A and the pair of discharge rollers 17B, thereby conveying the recording paper P fed by the paper feeding unit 14 along the conveying path T toward the image forming unit 12 and the discharge tray 8.

[0023] The HDD 18 stores various data including image data generated by the image reading unit 11 and various computer programs for realizing the operation of the image forming apparatus 1. The HDD 18 corresponds to the "storage unit" in the claims.

[0024] In this embodiment, the HDD 18 stores a first registration program for executing a first voice data registration process and a first playback program for executing a first voice data playback process. The HDD 18 also stores authentication information indicating a combination of a user ID and a password for each user.

[0025] The image processing unit 19 performs image processing as necessary on the image data generated by the image reading unit 11. The image memory 20 includes an area for temporarily storing the image data generated by the image reading unit 11.

[0026] The input unit 21 includes a microphone (hereinafter simply referred to as "mic") 21A and an A / D conversion circuit. The microphone 21A is provided on the surface of the housing of the operation unit 16. The user's voice is input to the microphone 21A. The A / D conversion circuit converts an analog signal representing the voice input to the microphone 21A into a digital signal.

[0027] The output unit 22 includes a speaker 22A, a D / A conversion circuit, and an amplifier. The speaker 22A is provided on the surface of the housing of the operation unit 16. The D / A conversion circuit converts a digital signal represented by audio data into an analog signal. The speaker 22A outputs audio amplified by the amplifier based on the converted analog signal.

[0028] The facsimile communication unit 23 connects to a public line and transmits and receives image data via the public line. The communication unit 24 includes a communication module such as a LAN (Local Area Network) board. The control unit 10 communicates data with an external device such as a PC (Personal Computer) 25 connected via the network via the communication unit 24.

[0029] A power supply is connected to each part of the image forming apparatus 1. When the power supply is turned on by a user, power is supplied to each part of the image forming apparatus 1 from the power supply.

[0030] In this embodiment, the control unit 10 operates in accordance with the first registration program, thereby generating unique identification information in accordance with predetermined rules as the image forming unit 12 generates a printed material, and executes a first voice data registration process in which the user's voice is input into the input unit 21, and the generated identification information is associated with voice data indicating the voice input into the input unit 21 and stored in a predetermined storage location.

[0031] The control unit 10 also operates in accordance with the first playback program, and when it receives identification information via the operation unit 16, it accesses a predetermined storage location and executes a first audio data playback process that causes the output unit 22 to output the audio indicated by the audio data corresponding to the received identification information.

[0032] [Operation] (1) Authentication process The operation of image forming apparatus 1 when performing authentication processing will be described with reference to FIG.

[0033] When the image forming apparatus 1 is powered on, the control unit 10 causes the display unit 15 to display the authentication screen 30 shown in FIG. 3. It is assumed that the user inputs a user ID and password using the touch panel 16B. The control unit 10 causes the display unit 15 to display the input user ID in box 32 and the input password in box 34. It is assumed that the user confirms the displayed user ID and password and touches the OK button 36.

[0034] When the OK button 36 is touched, the control unit 10 determines whether the input user ID and password match any of the authentication information stored in the HDD 18. In this case, the control unit 10 determines that the input user ID and password match the authentication information, and permits login to the image forming apparatus 1. Note that if the control unit 10 determines that the input user ID and password do not match the authentication information, it does not permit login to the image forming apparatus 1 and causes the display unit 15 to display an error message, for example, "Login not permitted."

[0035] (2) First voice data registration process 4 to 8, the operation of the image forming apparatus 1 when the first voice data registration process is executed will be described below. In this embodiment, it is assumed that the control unit 10 sets the HDD 18 as a predetermined storage destination for registering voice data.

[0036] When the control unit 10 permits the user to log in to the image forming apparatus 1, the control unit 10 causes the display unit 15 to display a job selection screen 40 shown in Fig. 4. It is assumed that the user touches a registration button 42. When the registration button 42 is touched, the control unit 10 starts a first registration program stored in the HDD 18, thereby starting the execution of the first voice data registration process shown in Fig. 5.

[0037] 6 on the display unit 15 (step S101). After the process of step S101, the control unit 10 repeats the process of determining that a print start instruction has not been input until the start key 16A is pressed (NO in step S102). In this situation, it is assumed that the user places multiple documents on the document platen of the image reading unit 11 and presses the start key 16A.

[0038] Control unit 10 determines that a print start instruction has been input (YES in step S102) and causes image reading unit 11 to convey multiple documents one by one using document conveying unit 6 and to read the image of each document using the scanner, thereby generating image data representing the document image (step S103). After processing step S103, control unit 10 generates unique identification information according to a predetermined rule (step S104). In this case, control unit 10 generates the unique identification information using information that is uniquely determined depending on the content of the image data generated by image reading unit 11 (for example, a hash value derived from the image data).

[0039] The control unit 10 may generate unique identification information using not only the image data but also at least one of the following information: an identifier of the image forming device 1 (such as a serial number or MAC address), the print date and time, an identifier of the user (printer) (such as a user ID or employee number), and an attribute value of the print job (such as size, color or monochrome, or total number of dots). In this case, the control unit 10 combines at least a portion of at least one of the above pieces of information to generate unique identification information.

[0040] After the process of step S104, the control unit 10 causes the display unit 15 to display the voice input screen 70 shown in Fig. 7 (step S105). After the process of step S105, the control unit 10 repeats the process of determining that a recording start instruction has not been input until the start button 72 is touched (NO in step S106). In this situation, it is assumed that the user has touched the start button 72.

[0041] The control unit 10 determines that an instruction to start recording has been input (YES in step S106), turns on the microphone 21A, and starts recording the audio (step S107). After the process of step S107, the control unit 10 repeats the process of determining that an instruction to stop recording has not been input (NO in step S108) until the stop button 74 is touched.

[0042] In this situation, the user inputs a voice into microphone 21A indicating a comment about the document image, such as "Please submit it by the end of this week." The A / D conversion circuit of input unit 21 converts the analog signal indicating the voice input into microphone 21A into a digital signal. Control unit 10 performs noise removal processing and band adjustment processing on the digital signal output from the A / D conversion circuit, and then temporarily stores the digital signal in RAM as RAW data.

[0043] Assume that after inputting the voice, the user touches stop button 74. Control unit 10 determines that a command to stop recording has been input (YES in step S108), turns off microphone 21A, and ends voice recording (step S109). At this time, control unit 10 performs encoding processing on the RAW data temporarily stored in RAM, and converts the RAW data into voice data in a predetermined data file format (e.g., voice data compressed in accordance with MP3 (Moving Picture Coding Experts Group-1 Audio Layer 3)). Note that control unit 10 may encrypt the generated voice data if it has previously received an encryption command from the user via operation unit 16.

[0044] After processing step S109, the control unit 10 generates a one-time password (step S110). In this embodiment, the authentication method for the generated one-time password is preferably counter-based. After processing step S110, the control unit 10 associates the voice data, unique identification information, and one-time password, and stores them in a predetermined storage location (in this case, HDD 18) (step S111). Here, the method of association is not particularly limited as long as it allows the voice data to be identified by the identification information. For example, association in a table format that allows reference on a database, or electronic association using a certain sealing process or compression / encryption process can be used.

[0045] After the process of step S111, the control unit 10 causes the display unit 15 to display the destination input screen 80 shown in FIG. 8 (step S112). After the process of step S112, the control unit 10 repeats the process of determining that an instruction to confirm the destination has not been input until the OK button 84 is touched (NO in step S113). In this situation, it is assumed that the user has input a destination address using the touch panel 16B. The control unit 10 causes the display unit 15 to display the input address in the box 82.

[0046] It is assumed that the user confirms the displayed address and touches the OK button 84. The control unit 10 determines that an instruction to confirm the destination has been input (YES in step S113), and sends an email indicating the presence of the voice data to the input destination address via the communication unit 24 (step S114). At this time, the control unit 10 generates an email indicating the generated unique identification information and one-time password. Note that the control unit 10 may generate an email that further indicates attribute information of the printed matter (title, number of pages, sender name, date and time of transmission, etc.) in addition to the identification information and one-time password.

[0047] After processing step S114, control unit 10 causes image forming unit 12 and the like to form an original image indicated by the image data generated by image reading unit 11 on recording paper P to generate a printed matter (step S115). After processing step S115, control unit 10 ends execution of the first voice data registration process. After the first voice data registration process is completed, the user delivers the printed matter generated by image forming device 1 to another user, for example, by internal mail.

[0048] (3) First audio data playback process Hereinafter, the operation of the image forming apparatus 1 when the first audio data reproduction process is executed will be described with reference to FIGS.

[0049] The other user who receives the printed material delivered as described above checks the email addressed to him / her and becomes aware of the presence of the voice data addressed to him / her. The other user then performs authentication processing using the user ID and password of the other user in the same manner as described above. When the control unit 10 permits the user to log in to the image forming apparatus 1, it causes the display unit 15 to display the job selection screen 40 shown in FIG. 4.

[0050] It is assumed that another user has touched the play button 44. When the play button 44 is touched, the control unit 10 starts the first playback program stored in the HDD 18, thereby starting the execution of the first audio data playback process shown in Fig. 9. The control unit 10 first causes the display unit 15 to display the identification information input screen 101 shown in Fig. 10 (step S201). After the process of step S201, the control unit 10 repeats the process of determining that an instruction to confirm the input information has not been input until the OK button 104 is touched (NO in step S202).

[0051] In this situation, it is assumed that the other user inputs the identification information indicated by the email using touch panel 16B. Control unit 10 acquires the input identification information and causes display unit 15 to display it in box 102. It is assumed that the user confirms the displayed identification information and touches OK button 104. Control unit 10 determines that a confirmation instruction has been input (YES in step S202), and determines whether or not the audio data corresponding to the acquired identification information is stored in a predetermined storage destination (in this case, HDD 18) (step S203).

[0052] In this case, control unit 10 determines that the corresponding voice data is stored (YES in step S203), and causes display unit 15 to display password entry screen 110 shown in Fig. 11 (step S204). After the process of step S204, control unit 10 repeats the process of determining that an instruction to confirm the password has not been input until OK button 114 is touched (NO in step S205). In this situation, it is assumed that the other user has used touch panel 16B to enter the one-time password indicated in the email.

[0053] The control unit 10 causes the display unit 15 to display the input one-time password in the box 112. It is assumed that the user confirms the displayed one-time password and touches the OK button 114. The control unit 10 determines that a confirmation instruction has been input (YES in step S205), and determines whether the input password matches the one-time password stored in a predetermined storage location (in this case, HDD 18) in association with the voice data corresponding to the input identification information (step S206).

[0054] In this case, control unit 10 determines that the passwords match (YES in step S206) and causes display unit 15 to display playback instruction input screen 120 shown in Fig. 12 (step S207). After the process of step S207, control unit 10 repeats the process of determining that a playback instruction has not been input until playback button 122 is touched (NO in step S208). In this situation, it is assumed that another user has touched playback button 122.

[0055] The control unit 10 determines that a playback instruction has been input (YES in step S208), and causes the output unit 22 to output the sound indicated by the sound data corresponding to the input identification information (step S209). At this time, if the sound data is encrypted, the control unit 10 decrypts the sound data using a previously acquired decryption key, and then causes the output unit 22 to output the sound.

[0056] After the process of step S209, the control unit 10 deletes the reproduced voice data and the identification information and one-time password corresponding to the voice data from a predetermined storage location (in this case, the HDD 18) (step S210). After the process of step S210, the control unit 10 ends the execution of the first voice data reproduction process.

[0057] If the control unit 10 determines that the audio data corresponding to the input identification information is not stored in the predetermined storage location (NO in step S203) or that the passwords do not match (NO in step S206), the control unit 10 causes the display unit 15 to display an error message such as "Audio cannot be played" (step S211). After the process of step S211, the control unit 10 ends the execution of the first audio data playback process.

[0058] According to the above embodiment, the control unit 10 generates unique identification information according to predetermined rules as the image forming unit 12 generates a printed matter, and causes the input unit 21 to input the user's voice, and associates the generated identification information with voice data indicating the voice input to the input unit 21 and stores them in a predetermined storage location.

[0059] This allows the user who created the printed material (printer) to save the audio associated with the generated printout, and other users who acquire the printout can use the identification information to identify and acquire the audio associated with the generated printout. This makes it possible to communicate verbally using spoken language, which cannot be obtained by simply exchanging printed materials. Therefore, even in a remote work environment, appropriate and smooth communication can be achieved using audio as well as paper media.

[0060] Furthermore, according to the above embodiment, the control unit 10 generates identification information according to a predetermined rule using information that is uniquely determined depending on the content of the image data, thereby easily generating unique identification information for each printed matter.

[0061] Furthermore, according to the above embodiment, the control unit 10 presets the HDD 18 as a predetermined storage destination, thereby enabling the identification information and voice data to be quickly registered.

[0062] Furthermore, according to the above embodiment, the control unit 10 receives an email destination address via the operation unit 16, and sends an email indicating the presence of voice data to the destination address via the communication unit 24. This allows other users who receive the email to easily recognize the presence of voice data.

[0063] Furthermore, according to the above embodiment, the control unit 10 generates an e-mail so that the e-mail includes identification information, which allows other users who receive the e-mail to easily obtain the identification information required to obtain the voice data.

[0064] Furthermore, according to the above embodiment, when the control unit 10 receives identification information via the operation unit 16, it accesses a predetermined storage location and causes the output unit 22 to output a sound represented by the audio data corresponding to the received identification information. This allows other users to play back the sound associated with the generated printed matter by simply inputting the identification information into the operation unit 16. This further improves convenience for other users.

[0065] Furthermore, according to the above embodiment, the control unit 10 deletes the audio represented by the audio data from a predetermined storage location when it causes the output unit 22 to output the audio, thereby enabling the audio data to be managed with high security.

[0066] (First Modification) In the above embodiment, the control unit 10 notifies other users of the existence of the audio data by email, but the present invention is not limited to such an embodiment. Hereinafter, in a first modification of the above embodiment, a configuration different from the above embodiment will be described, and the same configuration will not be described repeatedly.

[0067] In the first modified example, instead of sending an email, the control unit 10 notifies other users of the existence of the voice data by causing the image forming unit 12 or the like to form a message indicating the existence of the voice data together with the document image on the recording paper P. This allows other users who obtain the printed matter on which the document image and message are printed to easily recognize the existence of the voice data.

[0068] The control unit 10 also forms information indicating the identification information and one-time password, along with a message indicating the presence of audio data, on the recording paper P. This allows other users who obtain the printed matter to easily obtain the information necessary to play back the audio data.

[0069] (Second Modification) In the above embodiment, the control unit 10 acquires the identification information via the operation unit 16, but the present invention is not limited to such an embodiment. In a second modification of the above embodiment, the control unit 10 acquires the identification information by generating the identification information according to a predetermined rule using information that is uniquely determined depending on the content of the image data generated by the image reading unit 11. Below, in the second modification, a description will be given of configurations that are different from the above embodiment, and the same configurations will not be repeated.

[0070] In the second modification, the control unit 10 generates an e-mail indicating the one-time password and the storage location information without indicating the identification information in step S114 of the first voice data registration process. Therefore, other users do not obtain the identification information through the e-mail.

[0071] In the second modification, HDD 18 stores a second playback program for executing a second audio data playback process instead of the first playback program. Control unit 10 operates in accordance with the second playback program to generate identification information according to a predetermined rule using information uniquely determined according to the content of the image data generated by image reading unit 11, access a predetermined storage location, and execute a second audio data playback process that causes output unit 22 to output the audio indicated by the audio data corresponding to the generated identification information.

[0072] 13 and 14, the operation of the image forming device 1 when the second audio data reproduction process is executed will be described below. Regarding the second audio data reproduction process, the process that is different from the first audio data reproduction process shown in Fig. 19 will be described below, and the same process will not be described repeatedly. Also in Fig. 13, only the process that is different from the first audio data reproduction process shown in Fig. 19 will be shown.

[0073] It is assumed that another user performs authentication processing in the same manner as above and touches the play button 44 on the job selection screen 40 shown in Fig. 4. When the play button 44 is touched, the control unit 10 starts the second playback program stored in the HDD 18, thereby starting the execution of the second audio data playback process shown in Fig. 13. The control unit 10 first causes the display unit 15 to display the identification information input screen 140 shown in Fig. 14 (step S300).

[0074] After the process of step S300, control unit 10 repeats the process of determining that a reading start instruction has not been input until start key 16A is pressed (NO in step S301). In this situation, it is assumed that the user places multiple documents on the document platen of image reading unit 11 and presses start key 16A. Control unit 10 determines that a reading start instruction has been input (YES in step S301), and causes image reading unit 11 to transport the multiple documents one by one using document transport unit 6 and to read the image of each document using the scanner, thereby generating image data representing the document image (step S302).

[0075] After the process of step S302, the control unit 10 acquires the identification information by generating unique identification information according to a predetermined rule using information that is uniquely determined depending on the content of the image data generated by the image reading unit 11 (step S303), in the same manner as described above. After the process of step S303, the control unit 10 executes the processes from step S203 onward in the first audio data reproduction process shown in FIG.

[0076] According to the second modification, the control unit 10 acquires the identification information by generating the identification information according to a predetermined rule using information that is uniquely determined according to the content of the image data generated by the image reading unit 11. This allows other users to play back the generated audio related to the printed material by simply having the image reading unit 11 read the printed material they have obtained. This further improves convenience for other users.

[0077] (Third Modification) In the above embodiment, the control unit 10 stores the audio data in the HDD 18, but the present invention is not limited to such an embodiment. In a third modification of the above embodiment, the control unit 10 stores the audio data in cloud storage on a network. In the following, in the third modification, a configuration different from the above embodiment will be described, and the same configuration will not be described repeatedly.

[0078] In the third modified example, the control unit 10 sets a cloud storage on a network that can be communicated with via the communication unit 24 as a predetermined storage destination for registering voice data, instead of the HDD 18. Therefore, in step S111 of the first voice data registration process, the control unit 10 transmits a package file that associates the voice data, unique identification information, and one-time password to a cloud server on the network via the communication unit 24.

[0079] At this time, the control unit 10 cooperates with the cloud server by API control using, for example, authentication by an API (Application Programming Interface) key or OAuth authentication. The cloud server associates the voice data, unique identification information, and one-time password contained in the received package file, and stores them in a predetermined cloud storage on the network.

[0080] Furthermore, in the third modified example, in step S203 of the first audio data reproduction process, control unit 10 transmits the input identification information to a predetermined cloud server via communication unit 24. The cloud server determines whether audio data corresponding to the received identification information is stored in cloud storage, and transmits the determination result to image forming apparatus 1. If the determination result received via communication unit 24 indicates the presence of audio data, control unit 10 determines that the corresponding audio data is stored (YES in step S203). On the other hand, if the determination result indicates the absence of audio data, control unit 10 determines that the corresponding audio data is not stored (NO in step S203).

[0081] In step S206 of the first audio data reproduction process, control unit 10 also transmits the input password to a predetermined cloud server via communication unit 24. The cloud server determines whether the received password matches a one-time password stored in cloud storage in association with the previously received identification information, and transmits the determination result to image forming apparatus 1. If the determination result received via communication unit 24 indicates that the passwords match, control unit 10 determines that the passwords match (YES in step S206). On the other hand, if the determination result indicates that the passwords do not match, control unit 10 determines that the passwords do not match (NO in step S206).

[0082] In step S209 of the first audio data reproduction process, control unit 10 also transmits a request to transmit audio data to a predetermined cloud server via communication unit 24. The cloud server transmits audio data corresponding to the previously received identification information to image forming apparatus 1. Control unit 10 causes output unit 22 to output the audio indicated by the received audio data.

[0083] Furthermore, in step S210 of the first audio data playback process, the control unit 10 deletes the received audio data from the storage location (e.g., RAM) of the image forming device 1, and transmits an instruction to delete the audio data to a predetermined cloud server via the communication unit 24. Upon receiving the deletion instruction, the cloud server deletes the audio data transmitted to the image forming device 1 and the identification information and one-time password corresponding to the audio data from the cloud storage.

[0084] According to the third modification, the control unit 10 pre-sets, as a predetermined storage destination, a cloud storage on a network that can communicate via the communication unit 24. This not only enables the identification information and voice data to be quickly registered, but also makes it easy to ensure the storage capacity of the voice data.

[0085] (Fourth Modification) In the above embodiment, the other user plays back the sound indicated by the sound data using the same image forming device 1 as the image forming device 1 that registered the sound data, but the present invention is not limited to such an embodiment. In a fourth variation of the above embodiment, the other user plays back the sound indicated by the sound data using the image forming device 1 or another external device connected to a predetermined cloud server via a network.

[0086] The other external device may be, for example, another image forming device equipped with an output unit, a PC, or a smartphone. This further improves convenience for the user and other users, for example, when the user and other users work in a large office where multiple image forming devices 1 are installed, or when the user and other users cannot use the same image forming device 1 due to remote work.

[0087] In the fourth modification, in step S114 of the first voice data registration process, the control unit 10 preferably generates an e-mail indicating storage destination information (in this case, a URI (Uniform Resource Identifier) ​​or URL (Uniform Resource Locator)) indicating a predetermined storage destination in addition to the generated unique identification information and one-time password. Alternatively, the control unit 10 preferably causes the image forming unit 12 or the like to form a code, such as a two-dimensional code for accessing the predetermined storage destination or the voice data directly, on the recording paper P. This allows other users to easily access the storage destination of the voice data and play the voice by clicking on the storage destination information indicated in the e-mail or reading the code using a code reader provided in an external device.

[0088] (Other variations) In the above embodiment, the control unit 10 plays back the audio when a playback command is input, but the present invention is not limited to such an embodiment. For example, the control unit 10 may automatically play back the audio when it determines that the password matches, without waiting for the input of a playback command. The control unit 10 may also preset the timing for playing back these audios according to a user command received via the operation unit 16.

[0089] In the above embodiment, the control unit 10 generates a one-time password, but the present invention is not limited to such an embodiment. For example, the control unit 10 may generate a one-time password only when an instruction to generate a one-time password has been received in advance via the operation unit 16. Note that generating a one-time password allows for more secure management of voice data.

[0090] In the above embodiment, the control unit 10 deletes the audio data from a predetermined storage location when it causes the output unit 22 to output audio, but the present invention is not limited to such an embodiment. For example, the control unit 10 may delete the audio data when it receives a deletion instruction from the user via the operation unit 16 of the image forming apparatus 1 or an external device, or may delete the audio data after a predetermined period of time has elapsed. Alternatively, the control unit 10 may maintain the storage of the audio data semi-permanently without deleting it.

[0091] In the above embodiment, the control unit 10 records the voice during the processing of the copy job (from reading the document image to printing), but the present invention is not limited to such an embodiment. For example, the control unit 10 may record the voice immediately before the copy job (before reading the document image), or until a predetermined period has elapsed from the end of the copy job (for example, until the next copy job starts or until logging out).

[0092] In the above embodiment, the image forming apparatus 1 itself includes the input unit 21, but the present invention is not limited to such an embodiment. For example, a smartphone that can communicate with the image forming apparatus 1 via a wired or wireless connection such as Bluetooth may function as the input unit 21. In this case, the control unit of the smartphone transmits audio data indicating audio input to a microphone included in the smartphone to the image forming apparatus 1.

[0093] Furthermore, in the above embodiment, the control unit 10 only reproduces the sound of the voice data through the output unit 22, but the present invention is not limited to such an embodiment. For example, the control unit 10 may convert the sound represented by the voice data into text using existing voice recognition technology and cause the image forming unit 12 or the like to form the text representing the sound on the recording paper P in a predetermined format. This allows other users to visually recognize the sound represented by the voice data, further improving convenience.

[0094] In the above embodiment, the image forming unit 12 etc. formed an image on recording paper P, but the present invention is not limited to such an embodiment. The image forming unit 12 etc. may form an image on other recording media, not limited to recording paper. Examples of other recording media include overhead projector (OHP) sheets.

[0095] In the third modified example, the control unit 10 transmits a package file to the cloud server each time the first voice data registration process is executed, but the present invention is not limited to such an embodiment. For example, the control unit 10 may transmit the package file to the cloud server in a batch when the amount of data to be transmitted reaches a certain amount.

[0096] The present invention is not limited to the above-described embodiment and modified examples, and various modifications are possible. For example, in the above-described embodiment, a color multifunction peripheral is used as the image forming apparatus, but this is merely an example, and other image forming apparatuses such as a monochrome multifunction peripheral, a copier, or a facsimile machine may also be used.

[0097] The configurations and processes of the above-described embodiment and modified examples shown with reference to FIGS. 1 to 14 are merely one embodiment of the present invention, and the present invention is not limited to these configurations and processes. [Explanation of symbols]

[0098] 1. Image forming device 10 Control Unit 11 Image reading unit 12 Image forming unit 16 Control section 18 HDD 21 Input section 22 Output section 24 Communications Department

Claims

1. an image forming unit that forms an image indicated by the image data on a recording sheet to generate a print; an input unit for inputting a user's voice; an image forming device comprising: a control unit that generates unique identification information according to predetermined rules as the image forming unit generates the printed matter, has the user input voice into the input unit, and associates the generated identification information with voice data indicating the voice input into the input unit and stores them in a predetermined storage location.

2. The image forming apparatus according to claim 1 , wherein the control unit generates the identification information in accordance with the predetermined rule using information that is uniquely determined depending on the content of the image data.

3. Further comprising a storage unit, The image forming apparatus according to claim 1 , wherein the control unit pre-sets the storage unit as the predetermined save destination.

4. Further comprising a communication unit for communicating with an external device via a network, The image forming apparatus according to claim 1 , wherein the control unit pre-sets a cloud storage on the network that can communicate via the communication unit as the predetermined storage destination.

5. an operation unit into which user instructions are input; a communication unit that communicates with an external device via a network, 2. The image forming apparatus according to claim 1, wherein the control unit receives an email destination address via the operation unit, and sends an email indicating the presence of the voice data to the destination address via the communication unit.

6. The image forming apparatus according to claim 1 , wherein the control unit generates the email so as to include the identification information.

7. 2. The image forming apparatus according to claim 1, wherein the control unit causes the image forming unit to form, on the recording paper, an image represented by the image data and a message indicating the presence of the audio data.

8. The image forming apparatus according to claim 1 , wherein the control unit causes the image forming unit to form, on the recording paper, an image indicated by the image data and a code for accessing the predetermined storage location.

9. an operation unit into which user instructions are input; an output unit that outputs audio, 2. The image forming apparatus according to claim 1, wherein when the control unit receives the identification information via the operation unit, the control unit accesses the predetermined storage location and causes the output unit to output a sound indicated by the sound data corresponding to the received identification information.

10. an image reading unit that reads an image of a document and generates image data; an output unit that outputs audio, 3. The image forming apparatus according to claim 2, wherein the control unit generates the identification information in accordance with the predetermined rules using information uniquely determined according to the content of the image data generated by the image reading unit, accesses the predetermined storage location, and causes the output unit to output a sound indicated by the sound data corresponding to the generated identification information.

11. 11. The image forming apparatus according to claim 9, wherein the control unit deletes the audio data from the predetermined storage location after causing the output unit to output the audio.

Citation Information

Patent Citations

  • Image information preserving device

    JP2003122607A

  • Information recording medium, playback device and method thereof

    JP2007501485A

  • Voice processor, and program

    JP2008085486A

  • Information processing system

    JP2018148401A